Files
msd-core/tests/progress-forensic.test.cjs
Tom Boucher ff4a57b78c chore(#1671): migrate the remaining 13 LARGE/XL workflows to the fragment model — Phase 6.3 (#3030)
* chore(#2994): fragmentize progress.md forensic audit onto the fragment model

Extract the --forensic-gated forensic_audit step to
workflows/progress/steps/forensic-audit.md behind a section marker, and
repair progress.md's init line to forward --forensic so the atom is
actually true in production rather than only under direct CLI tests.

progress.md shrinks 32630 -> 27207 bytes.

Co-Authored-By: Claude Opus 5 (1M context) <noreply@anthropic.com>

* chore(#2994): fragmentize the four manifest-wired workflows

new-project, quick, new-milestone and progress each already had a
dedicated cmdInit* entry point but zero marked sections. Extract nine
gated bodies to workflows/<wf>/steps/ behind section markers and repair
each init line to forward its flags.

Fold --full into the discuss/research/validate facts inside cmdInitQuick
so the when= grammar never sees an OR, per the chunked-mode precedent.

Fixes found while working, per the no-defer rule:
- cmdInitProgress passed no phase info to buildSectionManifestField, so
  state:phase-mvp-mode was permanently false — an atom in the vocabulary
  whose fact could never be computed.
- the quick init router folded flag tokens into the free-text
  description, which the new forwarding would have corrupted.
- a #2508 dispatch note was nested inside quick.md's Agent(prompt=)
  fence, leaking orchestrator guidance into the subagent prompt.
- progress.md had a 3-vs-4 backtick outer-fence imbalance.

Co-Authored-By: Claude Opus 5 (1M context) <noreply@anthropic.com>

* chore(#2994): fragmentize verify-work.md and admit state:ui-phase-active

Wire cmdInitVerifyWork to buildSectionManifestField — it was a dedicated
entry point that never emitted a manifest — and mark two sections.

state:ui-phase-active folds (plan:pre hooks include an active ui step) OR
(the phase dir holds a *-UI-SPEC.md) into one boolean in init.cts, so the
grammar still sees a single operator-free atom. The inner Playwright-MCP
check stays as prose inside the fragment: it is live session state and no
init seam can precompute it.

The MVP false-branch note is a real fallback, not redundant prose, so it
sits outside the marker — gating it away would delete the text needed
precisely when MVP mode is off.

Co-Authored-By: Claude Opus 5 (1M context) <noreply@anthropic.com>

* test(#2994): follow moved workflow content in drift guards

Retarget every guard that asserted on content this branch moved into
workflows/<wf>/steps/, mirroring 815b3d897. Each retargeted assertion was
verified to still fail when its step file is blanked, so none was
weakened into vacuity.

Three assertions in verify-mvp-uat were genuinely red. Three more were
worse than red — passing for the wrong reason:
- quick-commit-boundary and worktree-cleanup anchored on indexOf('Step
  5.6'), which matched a later cross-reference and sliced 16069 chars
  that coincidentally held the asserted substrings. Replaced with an
  expandWorkflowSections helper that splices step content back in place.
- phase6-review-capabilities lost its end boundary and widened to EOF.
- playwright-ui-verify matched 'UI' in an unrelated bullet and 'fall
  back' in a subagent-dispatch line after the real content moved.

Co-Authored-By: Claude Opus 5 (1M context) <noreply@anthropic.com>

* chore(#2994): fragmentize code-review and complete-milestone, admit three atoms

Add dedicated cmdInitCodeReview and cmdInitCompleteMilestone entry points
alongside the shared generic ones rather than modifying them — init.phase-op
and init.manager carry a CRITICAL blast radius (179 dependents, 24
processes) and stay byte-identical for their other callers.

Admit flag:--fix, state:fallow-enabled and state:git-create-tag, each with
a consuming section and a fact its own entry point computes.

Both sections had the resolver-in-body hazard: the fallow config-gate and
the git.create_tag check each sat inside the very block being gated, so
gating would have disabled the resolver that decides the gate. Both are
hoisted into init and the bodies now consume the resolved fact.

Co-Authored-By: Claude Opus 5 (1M context) <noreply@anthropic.com>

* test(#2994): retarget code-review and milestone drift guards, fix two red tests

Retarget guards that asserted on content moved into steps/, proving
non-vacuity by blanking each step file and confirming failure.

Also fixes two genuinely red tests found while working, per the no-defer
rule:
- workflow-fragments' frozen-vocabulary lock was missing
  state:ui-phase-active, so commit 7ef7f8336 shipped red. Lint and build
  both passed over it, which is why neither is sufficient verification.
- code-review's quick.md capability-hook assertion carried a stale
  delimiter after the 18ff35d20 extraction.

Co-Authored-By: Claude Opus 5 (1M context) <noreply@anthropic.com>

* chore(#2994): fragmentize autonomous.md and admit state:plan-strategy-converge

Five sections share one atom, the pattern plan-phase already uses for
flag:--research-phase. The atom folds --converge OR --cross-ai into a
single boolean in cmdInitAutonomous so the grammar stays operator-free.

cmdInitAutonomous is additive; init.milestone-op, init.manager and
init.phase-op are untouched and still consumed. The $PLAN_STRATEGY bash
resolver is deliberately retained — ungated local-planning bullets still
read it, so the init-side fact supplements it rather than replacing it.

converge-fail-fast required splitting one bash fence so the always-run
CONVERGENCE_ARGS construction stays outside the marker. All three
flag-absent fallbacks were left outside their markers.

Co-Authored-By: Claude Opus 5 (1M context) <noreply@anthropic.com>

* chore(#2994): fragmentize review and discuss-phase-assumptions

Admit state:reviewer-instances-configured (two peripheral notes share it;
the core reviewer-lane dispatch stays unmarked — it is the workflow's
primary always-evaluated logic, not an optional branch) and
state:auto-advance-active, which folds --auto OR two config keys into one
boolean so the grammar stays operator-free.

discuss-phase-assumptions was the highest-risk edit in this PR. Its
auto_advance step is a full if/elif/else; gating it whole would have
deleted the flag-absent fallback needed exactly when --auto is off. Split
verified exact: resolvers 636-651 and the 'End here' fallback 668-669 both
stay outside the marker; only 653-667 is gated.

Adds emitted-drift acks for the two files that grew — review.md (+55 B)
and autonomous.md (+737 B from 80799211c, which had none and would have
red-gated the push.

Co-Authored-By: Claude Opus 5 (1M context) <noreply@anthropic.com>

* chore(#2994): fragmentize docs-update, update, transition and new-milestone Part A

Completes the 13-workflow rollout. Three of these had no init call at all
and gained a dedicated entry point plus their first gsd_run query line.

Admits state:is-monorepo and adds state:next-channel, state:workstream-active
and state:flat-mode. Vocabulary 26 -> 30 atoms.

Part A of new-milestone applies when NO workstream is active — the negation
of state:workstream-active. Rather than teach the grammar negation, which is
the Greenspun drift the frozen list exists to prevent, it gets a separate
positively-phrased atom whose fact is the inverse. Part B, which always runs,
stays outside the marker.

flag:--verify-only is deliberately NOT admitted: docs-update has no
contiguous purely-additive region for it, and an atom without a consuming
section is dead vocabulary. Evidence recorded in the slice report.

update.md reuses its existing resolved $GSD_TOOLS rather than prepending the
canonical preamble, which would have clobbered it.

Co-Authored-By: Claude Opus 5 (1M context) <noreply@anthropic.com>

* fix(#2994): stop automated-ui-verification re-resolving its own gate, retire dead vocabulary

Two defects the new tests caught.

The automated-ui-verification step re-ran gsd_run loop render-hooks and
recomputed UI_PHASE_ACTIVE inside a body that is only read when that fact
is already true — the circular self-disabling pattern this design forbids,
introduced by 3c654b168. cmdInitVerifyWork now exposes ui_phase_active and
the step consumes it. Its launcher preamble goes too: no gsd_run remains.
The Playwright-MCP check stays as prose — that is live session state.

Dead vocabulary predating this PR: flag:--full and state:needs-codebase-map
were admitted with a gate-1 claim that never materialized. flag:--full is
removed, redundant once quick folds it into discuss/research/validate.
state:needs-codebase-map gets the real consumer it always lacked, gating
new-project's codebase-map offer. Vocabulary 30 -> 29, and no atom is now
without a consuming section.

Co-Authored-By: Claude Opus 5 (1M context) <noreply@anthropic.com>

* test(#2994): add the atom-admission, inversion and resolver-hoist gates

The two existing parity guards prove vocabulary/predicate symmetry but
never that a fact is computed — an atom no cmdInit* assembles evaluates
false forever. These close that hole:

- per-atom satisfiability for all 29 atoms, plus an anti-vacuity assertion
  so the loop cannot silently cover zero atoms
- dead-vocabulary check against the shipped manifest
- inversion guard: the flag-absent fallbacks in discuss-phase-assumptions
  and verify-work must stay outside their markers
- data-driven resolver-hoist guard over the shipped manifest, so a future
  extraction cannot reintroduce the circular class
- compound-fold coverage (--full, --cross-ai, --rc, config-only --auto)
- null-vs-[] degraded/computed distinction, and flag value shapes

Also repairs the frozen-vocabulary lock, which was stale and red for the
seven atoms earlier commits on this branch shipped.

Co-Authored-By: Claude Opus 5 (1M context) <noreply@anthropic.com>

* docs(#2994): add changeset for the fragment-model rollout

Co-Authored-By: Claude Opus 5 (1M context) <noreply@anthropic.com>

* test(#2994): cite the issue on the two new allow-test-rule exemptions

ADR-456 requires an issue ref on the same line as the annotation.

Co-Authored-By: Claude Opus 5 (1M context) <noreply@anthropic.com>

* docs(#2994): correct the atom-count claims after retiring flag:--full

The vocabulary doc comments still said 30 entries; it is 29 since
flag:--full was removed as dead vocabulary.

Co-Authored-By: Claude Opus 5 (1M context) <noreply@anthropic.com>

* fix(#2994): dedupe the phase-fallback block and harden --ws parsing

Review findings.

MAJOR: the three new init entry points each pasted a verbatim copy of the
guardedFindPhase/guardedGetRoadmapPhase fallback, taking the repo from four
copies to seven — DEFECT.GENERATIVE-FIX. Extracted applyRoadmapFallback and
folded six of the seven; each call site keeps its own field-set via a
closure. Duplication removed rather than papered over with a parity test.
cmdInitPhaseOp stays out: its fallback omits has_reviews, so it is not a
byte-identical copy, and it is CRITICAL-radius.

LOW, pre-existing: GSD_WS captured [^[:space:]]+ and expands unquoted, so a
workstream name holding glob metacharacters would expand against the
filesystem. Narrowed to [A-Za-z0-9._-]+. The unquoted expansion is kept —
it must word-split into two args and vanish when empty.

Also restores the vocabulary ordering convention, and fixes a masked test
bug the mandated run surfaced: the flag-forwarding guard checked only the
first init line per workflow, but new-milestone has two, so a real failure
was reporting exit 0.

Co-Authored-By: Claude Opus 5 (1M context) <noreply@anthropic.com>

* fix(#2994): drop the stale new-milestone emitted-drift ack

new-milestone.md was acked for a +406 B growth measured against an
intermediate commit. Net against origin/next it SHRANK by 8 bytes, so
nothing needed the ack and it explained nothing — which the differential
attribution check reports as a stale acknowledgment, not a pass.

update.md's entry stays: it genuinely grew +703 B.

Co-Authored-By: Claude Opus 5 (1M context) <noreply@anthropic.com>

* fix(#2994): resolve the 15 failures from the full matrix run

All 15 were real and identical on both lanes.

REAL REGRESSION: autonomous.md hit 41479 chars against the #2196 guard's
40960 cap — a CHARS cap distinct from the LARGE tier byte cap, which the
five section stubs pushed it over. Extracted the 3a.5 UI Design Contract
body to references/; now 39968 chars, and the file nets -795 B vs base, so
its growth ack is deleted rather than left stale.

REAL DEFECT: docs referenced /gsd-transition, which is not a live
registered command. Reworded.

STALE FIXTURE: the emission byte-identity test hardcoded two marked
workflows; this branch legitimately marks fifteen. Fixture corrected — the
source was right.

The rest were drift guards over the eight workflows the earlier sweep did
not cover, retargeted at where the content now lives with non-vacuity
proven by blanking each step file and confirming failure. The GSD_WS
forwarding guard was checked as a possible real break and is not one: the
charclass narrowing is intact and forwarding works end to end.

Co-Authored-By: Claude Opus 5 (1M context) <noreply@anthropic.com>

* fix(#2994): drop the ack for a newly-added reference file

A new file's emitted ripple is attributable to the diff that adds it, so
the acknowledgment explained nothing and the differential check reports it
as stale. Removing the last entry removes the fragment — an empty one
signals nothing.

Co-Authored-By: Claude Opus 5 (1M context) <noreply@anthropic.com>

* fix(#2994): retarget the UI-contract guards and clear two transitive advisories

The §3a.5 extraction that brought autonomous.md under the #2196 char cap
moved its body to references/autonomous-ui-design-contract.md, so ten
guards in autonomous-ui-steps and check-ui-safety-gate were asserting it
against the host. Retargeted via a combined read, each proven non-vacuous
by blanking the reference file and confirming failure.

This class had already bitten twice on this branch because each sweep was
scoped to the workflows touched at that moment, so this one was
exhaustive: ~70 test files across all 13 workflows, zero further broken or
vacuous assertions found.

Also clears two high transitive advisories the matrix flagged on one lane
— fast-uri GHSA-7p8r-x3mc-p8w7 and three ip-address SSRF/trust-boundary
issues. Both pre-date this branch: package-lock.json was untouched until
now, so the production tree was byte-identical to the base. Lockfile-only,
package.json unchanged, verified against a real npm ci install.

Co-Authored-By: Claude Opus 5 (1M context) <noreply@anthropic.com>

* chore(#2994): backfill changeset pr number to 3030

---------

Co-authored-by: sim <sim@local>
Co-authored-by: Claude Opus 5 (1M context) <noreply@anthropic.com>
2026-08-03 19:59:58 -04:00

441 lines
18 KiB
JavaScript

// allow-test-rule: source-text-is-the-product
// Reads .md/.json/.yml product files whose deployed text IS what the
// runtime loads — testing text content tests the deployed contract.
/**
* Tests for --forensic flag on /gsd-progress (#2189)
*
* The --forensic flag appends a 6-check integrity audit after the standard
* progress report. Default behavior (no flag) is unchanged.
*/
const { test, describe } = require('node:test');
const assert = require('node:assert/strict');
const fs = require('fs');
const path = require('path');
// #2994 fragmentization moved the --forensic-gated forensic_audit step out of
// progress.md into gsd-core/workflows/progress/steps/forensic-audit.md behind
// a section marker. Tests that need the step BODY read that step file
// directly (it is the sole remaining source of the step's content); the
// step-presence check below is the only one that must also confirm the host
// still wires up the marker that reads it.
const FORENSIC_AUDIT_STEP_PATH = path.join(
__dirname, '..', 'gsd-core', 'workflows', 'progress', 'steps', 'forensic-audit.md'
);
describe('#2189: progress --forensic flag', () => {
test('progress command argument-hint includes --forensic', () => {
const command = fs.readFileSync(
path.join(__dirname, '..', 'commands', 'gsd', 'progress.md'), 'utf8'
);
assert.ok(command.includes('--forensic'), 'argument-hint should include --forensic');
});
test('progress workflow has a forensic_audit step', () => {
const workflow = fs.readFileSync(
path.join(__dirname, '..', 'gsd-core', 'workflows', 'progress.md'), 'utf8'
);
const step = fs.readFileSync(FORENSIC_AUDIT_STEP_PATH, 'utf8');
assert.ok(
workflow.includes('id="forensic-audit"'),
'progress.md must wire up the forensic-audit section marker'
);
assert.ok(
step.includes('<step name="forensic_audit">'),
'progress/steps/forensic-audit.md should have a forensic_audit step'
);
});
test('forensic_audit step is only triggered when --forensic is present', () => {
const forensicStep = fs.readFileSync(FORENSIC_AUDIT_STEP_PATH, 'utf8');
assert.ok(
forensicStep.includes('--forensic'),
'forensic_audit step should be gated on --forensic flag'
);
assert.ok(
forensicStep.includes('Skip') || forensicStep.includes('skip') || forensicStep.includes('exit'),
'forensic_audit step should skip when --forensic is not present'
);
});
test('forensic_audit step includes all 6 checks', () => {
const forensicStep = fs.readFileSync(FORENSIC_AUDIT_STEP_PATH, 'utf8');
// Check 1: STATE vs artifact consistency
assert.ok(
forensicStep.includes('STATE') && (forensicStep.includes('artifact') || forensicStep.includes('consistent')),
'forensic step should check STATE vs artifact consistency (check 1)'
);
// Check 2: Orphaned handoff files
assert.ok(
forensicStep.includes('HANDOFF') || forensicStep.includes('handoff'),
'forensic step should check for orphaned handoff files (check 2)'
);
// Check 3: Deferred scope drift
assert.ok(
forensicStep.includes('deferred') || forensicStep.includes('defer'),
'forensic step should check for deferred scope drift (check 3)'
);
// Check 4: Memory-flagged pending work
assert.ok(
forensicStep.includes('MEMORY') || forensicStep.includes('memory') || forensicStep.includes('pending'),
'forensic step should check memory-flagged pending work (check 4)'
);
// Check 5: Blocking todos
assert.ok(
forensicStep.includes('todo') || forensicStep.includes('Todo') || forensicStep.includes('TODO'),
'forensic step should check blocking operational todos (check 5)'
);
// Check 6: Uncommitted code
assert.ok(
forensicStep.includes('uncommitted') || forensicStep.includes('git status'),
'forensic step should check for uncommitted code (check 6)'
);
});
test('forensic_audit step produces a CLEAN or INTEGRITY ISSUE(S) FOUND verdict', () => {
const forensicStep = fs.readFileSync(FORENSIC_AUDIT_STEP_PATH, 'utf8');
assert.ok(
forensicStep.includes('CLEAN'),
'forensic step should produce a CLEAN verdict when all checks pass'
);
assert.ok(
forensicStep.includes('INTEGRITY ISSUE') || forensicStep.includes('integrity issue'),
'forensic step should surface INTEGRITY ISSUE when checks fail'
);
});
test('forensic_audit step does not change default progress behavior', () => {
// The forensic step must explicitly say default behavior is unchanged
const forensicStep = fs.readFileSync(FORENSIC_AUDIT_STEP_PATH, 'utf8');
assert.ok(
forensicStep.includes('unchanged') || forensicStep.includes('standard report'),
'forensic step should clarify that default behavior is unchanged'
);
});
test('COMMANDS.md documents --forensic flag for gsd-progress', () => {
const commands = fs.readFileSync(
path.join(__dirname, '..', 'docs', 'COMMANDS.md'), 'utf8'
);
assert.ok(
commands.includes('--forensic'),
'COMMANDS.md should document --forensic flag for gsd-progress'
);
});
});
/**
* Regression — issue #1107
*
* /gsd-progress reported a phase as complete and routed to the next phase even
* when its VERIFICATION.md ended `human_needed` / `gaps_found`, because routing
* derived completeness from plan/summary counts only and never consulted the
* `verification.status` query (built in #651). The fix adds a Step 1.7 consult
* and routing rows that send non-`passed` phases back to close the debt.
*/
describe('#1107: progress routing consults verification.status before reporting complete', () => {
function readWorkflow() {
return fs.readFileSync(
path.join(__dirname, '..', 'gsd-core', 'workflows', 'progress.md'), 'utf8'
);
}
test('workflow consults verification.status for the current phase', () => {
const workflow = readWorkflow();
assert.ok(
workflow.includes('verification.status'),
'progress workflow must query verification.status (the #651 seam)'
);
assert.ok(
workflow.includes('verification_status'),
'progress workflow must track a verification_status value for routing'
);
assert.ok(
workflow.includes('stale verification'),
'progress workflow must document that verification.status projects stale verification'
);
});
test('routing table has gaps_found and human_needed rows BEFORE the generic complete row', () => {
const workflow = readWorkflow();
const missingIdx = workflow.indexOf('verification_status = missing');
const unknownIdx = workflow.indexOf('verification_status = unknown');
const staleIdx = workflow.indexOf('verification_status = stale');
const gapsIdx = workflow.indexOf('verification_status = gaps_found');
const humanIdx = workflow.indexOf('verification_status = human_needed');
const completeIdx = workflow.indexOf('Phase complete (verification passed)');
assert.ok(missingIdx > -1, 'routing table must have a missing verification row');
assert.ok(unknownIdx > -1, 'routing table must have an unknown verification row');
assert.ok(staleIdx > -1, 'routing table must have a stale verification row');
assert.ok(gapsIdx > -1, 'routing table must have a gaps_found row');
assert.ok(humanIdx > -1, 'routing table must have a human_needed row');
assert.ok(completeIdx > -1, 'routing table must keep a generic complete row');
assert.ok(
missingIdx < completeIdx &&
unknownIdx < completeIdx &&
staleIdx < completeIdx &&
gapsIdx < completeIdx &&
humanIdx < completeIdx,
'verification rows must precede the generic "summaries = plans" complete row (first-match-wins)'
);
});
test('gaps_found routes to plan-phase --gaps (Route V.gaps)', () => {
const workflow = readWorkflow();
// Anchor on the definition heading (`**Route V.gaps:`), not the routing-table
// reference (`Go to **Route V.gaps**`).
assert.ok(workflow.includes('**Route V.gaps:'), 'must define a Route V.gaps section');
const route = workflow.slice(
workflow.indexOf('**Route V.gaps:'),
workflow.indexOf('**Route V.human:')
);
assert.ok(
route.includes('--gaps') && route.includes('plan-phase'),
'Route V.gaps must route to /gsd:plan-phase {phase} --gaps'
);
});
test('human_needed routes to verify-work (Route V.human)', () => {
const workflow = readWorkflow();
assert.ok(workflow.includes('**Route V.human:'), 'must define a Route V.human section');
const route = workflow.slice(
workflow.indexOf('**Route V.human:'),
workflow.indexOf('**Step 3', workflow.indexOf('**Route V.human:'))
);
assert.ok(
route.includes('verify-work'),
'Route V.human must route to /gsd:verify-work {phase}'
);
});
test('stale verification routes to verify-work (Route V.stale)', () => {
const workflow = readWorkflow();
assert.ok(workflow.includes('**Route V.stale:'), 'must define a Route V.stale section');
const route = workflow.slice(
workflow.indexOf('**Route V.stale:'),
workflow.indexOf('**Route V.gaps:')
);
assert.ok(
route.includes('verify-work'),
'Route V.stale must route to /gsd:verify-work {phase}'
);
});
test('missing and unknown verification do not route as complete', () => {
const workflow = readWorkflow();
assert.ok(
workflow.includes('Phase complete (verification passed)'),
'the generic complete row must only cover passed verification'
);
assert.ok(!workflow.includes('verification passed, missing, or n/a'),
'missing or unknown verification must not be documented as complete');
});
});
// ────────────────────────────────────────────────────────────────────────
// Folded from tests/bug-3418-progress-flag-routing.test.cjs — consolidation epic #1969 (B3 #1972)
// ────────────────────────────────────────────────────────────────────────
{
const { describe: __foldDescribe } = require('node:test');
__foldDescribe("folded:bug-3418-progress-flag-routing (consolidation epic #1969 B3 #1972)", () => {
// allow-test-rule: source-text-is-the-product (see #3418)
// The command markdown is loaded directly by runtime prompt assembly.
const { test, describe } = require('node:test');
const assert = require('node:assert/strict');
const fs = require('fs');
const path = require('path');
describe('#3418: /gsd-progress flag routing prompt contract', () => {
test('progress command surfaces raw arguments on a dedicated line before routing parse', () => {
const command = fs.readFileSync(
path.join(__dirname, '..', 'commands', 'gsd', 'progress.md'),
'utf8'
);
assert.ok(
command.includes('Arguments provided: "$ARGUMENTS"'),
'progress.md must surface $ARGUMENTS on a dedicated line for stable flag parsing'
);
});
test('progress command must not inline-substitute $ARGUMENTS into parse instruction text', () => {
const command = fs.readFileSync(
path.join(__dirname, '..', 'commands', 'gsd', 'progress.md'),
'utf8'
);
assert.ok(
!command.includes('Parse the first token of $ARGUMENTS:'),
'progress.md must keep parse instructions independent from argument interpolation'
);
});
});
});
}
// ────────────────────────────────────────────────────────────────────────
// Folded from tests/bug-2912-progress-context-authority.test.cjs — consolidation epic #1969 (B4 #1973)
// ────────────────────────────────────────────────────────────────────────
{
const { describe: __foldDescribe } = require('node:test');
__foldDescribe("folded:bug-2912-progress-context-authority (consolidation epic #1969 B4 #1973)", () => {
/**
* Tests for issue #2912 — /gsd-progress can use stale CLAUDE.md project block
* instead of GSD tracking files as authoritative source.
*
* Fix: the `report` step in gsd-core/workflows/progress.md must contain
* an explicit "context authority" directive establishing PROJECT.md, STATE.md,
* and ROADMAP.md as the authoritative sources for the progress report, and
* forbidding the use of CLAUDE.md `## Project` blocks as a source for any
* report field.
*
* These tests parse the workflow markdown structurally (locate the
* <step name="report"> ... </step> block, then locate the blockquote-style
* directive inside it). They do NOT use `.includes()` over the whole file.
*/
const { test, describe } = require('node:test');
const assert = require('node:assert/strict');
const fs = require('fs');
const path = require('path');
const WORKFLOW_PATH = path.join(
__dirname,
'..',
'gsd-core',
'workflows',
'progress.md'
);
/** Extract the body of a <step name="..."> ... </step> block by parsing tags. */
function extractStep(workflow, stepName) {
const openTag = `<step name="${stepName}">`;
const start = workflow.indexOf(openTag);
if (start === -1) return null;
const bodyStart = start + openTag.length;
// Find the matching </step> — workflow steps in this file do not nest.
const end = workflow.indexOf('</step>', bodyStart);
if (end === -1) return null;
return workflow.slice(bodyStart, end);
}
/**
* Extract contiguous markdown blockquote blocks from a chunk of markdown.
* A blockquote is a run of consecutive lines starting with '>' (after any
* leading whitespace). Returns the joined text of each blockquote with the
* leading '>' markers stripped.
*/
function extractBlockquotes(md) {
const lines = md.split(/\r?\n/);
const blocks = [];
let current = null;
for (const line of lines) {
const m = line.match(/^\s*>\s?(.*)$/);
if (m) {
if (current === null) current = [];
current.push(m[1]);
} else {
if (current !== null) {
blocks.push(current.join('\n'));
current = null;
}
}
}
if (current !== null) blocks.push(current.join('\n'));
return blocks;
}
describe('#2912: progress report step has explicit context-authority directive', () => {
test('progress.md workflow file exists and is readable', () => {
const stat = fs.statSync(WORKFLOW_PATH);
assert.ok(stat.isFile(), 'workflow file should exist');
});
test('progress.md has a <step name="report"> section', () => {
const workflow = fs.readFileSync(WORKFLOW_PATH, 'utf8');
const reportStep = extractStep(workflow, 'report');
assert.ok(reportStep, 'workflow should contain a report step');
assert.ok(reportStep.length > 0, 'report step body should not be empty');
});
test('report step contains a blockquote directive about context authority', () => {
const workflow = fs.readFileSync(WORKFLOW_PATH, 'utf8');
const reportStep = extractStep(workflow, 'report');
assert.ok(reportStep, 'report step must be present');
const blockquotes = extractBlockquotes(reportStep);
assert.ok(
blockquotes.length > 0,
'report step should contain at least one blockquote (the context-authority directive)'
);
const authorityBlock = blockquotes.find((b) => /context\s+authority/i.test(b));
assert.ok(
authorityBlock,
'report step should contain a blockquote whose text includes "Context authority"'
);
});
test('context-authority directive names PROJECT.md, STATE.md, and ROADMAP.md as authoritative', () => {
const workflow = fs.readFileSync(WORKFLOW_PATH, 'utf8');
const reportStep = extractStep(workflow, 'report');
assert.ok(reportStep, 'report step must exist');
const blockquotes = extractBlockquotes(reportStep);
const authorityBlock = blockquotes.find((b) => /context\s+authority/i.test(b));
assert.ok(authorityBlock, 'authority blockquote must exist');
assert.match(
authorityBlock,
/PROJECT\.md/,
'directive should name PROJECT.md as authoritative'
);
assert.match(
authorityBlock,
/STATE\.md/,
'directive should name STATE.md as authoritative'
);
assert.match(
authorityBlock,
/ROADMAP\.md/,
'directive should name ROADMAP.md as authoritative'
);
assert.match(
authorityBlock,
/authoritative/i,
'directive should describe these files as authoritative'
);
});
test('context-authority directive forbids using CLAUDE.md project block as a source', () => {
const workflow = fs.readFileSync(WORKFLOW_PATH, 'utf8');
const reportStep = extractStep(workflow, 'report');
assert.ok(reportStep, 'report step must exist');
const blockquotes = extractBlockquotes(reportStep);
const authorityBlock = blockquotes.find((b) => /context\s+authority/i.test(b));
assert.ok(authorityBlock, 'authority blockquote must exist');
assert.match(
authorityBlock,
/CLAUDE\.md/,
'directive should explicitly mention CLAUDE.md'
);
// Must explicitly forbid CLAUDE.md as a source — look for a NOT/do not directive
// co-located with the CLAUDE.md mention.
assert.match(
authorityBlock,
/(do\s+NOT|do\s+not|must\s+NOT|must\s+not|never)/i,
'directive should contain an explicit prohibition (do NOT / must not / never)'
);
assert.match(
authorityBlock,
/## Project/,
'directive should call out the CLAUDE.md "## Project" block specifically'
);
});
});
});
}