* chore(#2994): fragmentize progress.md forensic audit onto the fragment model Extract the --forensic-gated forensic_audit step to workflows/progress/steps/forensic-audit.md behind a section marker, and repair progress.md's init line to forward --forensic so the atom is actually true in production rather than only under direct CLI tests. progress.md shrinks 32630 -> 27207 bytes. Co-Authored-By: Claude Opus 5 (1M context) <noreply@anthropic.com> * chore(#2994): fragmentize the four manifest-wired workflows new-project, quick, new-milestone and progress each already had a dedicated cmdInit* entry point but zero marked sections. Extract nine gated bodies to workflows/<wf>/steps/ behind section markers and repair each init line to forward its flags. Fold --full into the discuss/research/validate facts inside cmdInitQuick so the when= grammar never sees an OR, per the chunked-mode precedent. Fixes found while working, per the no-defer rule: - cmdInitProgress passed no phase info to buildSectionManifestField, so state:phase-mvp-mode was permanently false — an atom in the vocabulary whose fact could never be computed. - the quick init router folded flag tokens into the free-text description, which the new forwarding would have corrupted. - a #2508 dispatch note was nested inside quick.md's Agent(prompt=) fence, leaking orchestrator guidance into the subagent prompt. - progress.md had a 3-vs-4 backtick outer-fence imbalance. Co-Authored-By: Claude Opus 5 (1M context) <noreply@anthropic.com> * chore(#2994): fragmentize verify-work.md and admit state:ui-phase-active Wire cmdInitVerifyWork to buildSectionManifestField — it was a dedicated entry point that never emitted a manifest — and mark two sections. state:ui-phase-active folds (plan:pre hooks include an active ui step) OR (the phase dir holds a *-UI-SPEC.md) into one boolean in init.cts, so the grammar still sees a single operator-free atom. The inner Playwright-MCP check stays as prose inside the fragment: it is live session state and no init seam can precompute it. The MVP false-branch note is a real fallback, not redundant prose, so it sits outside the marker — gating it away would delete the text needed precisely when MVP mode is off. Co-Authored-By: Claude Opus 5 (1M context) <noreply@anthropic.com> * test(#2994): follow moved workflow content in drift guards Retarget every guard that asserted on content this branch moved into workflows/<wf>/steps/, mirroring 815b3d897. Each retargeted assertion was verified to still fail when its step file is blanked, so none was weakened into vacuity. Three assertions in verify-mvp-uat were genuinely red. Three more were worse than red — passing for the wrong reason: - quick-commit-boundary and worktree-cleanup anchored on indexOf('Step 5.6'), which matched a later cross-reference and sliced 16069 chars that coincidentally held the asserted substrings. Replaced with an expandWorkflowSections helper that splices step content back in place. - phase6-review-capabilities lost its end boundary and widened to EOF. - playwright-ui-verify matched 'UI' in an unrelated bullet and 'fall back' in a subagent-dispatch line after the real content moved. Co-Authored-By: Claude Opus 5 (1M context) <noreply@anthropic.com> * chore(#2994): fragmentize code-review and complete-milestone, admit three atoms Add dedicated cmdInitCodeReview and cmdInitCompleteMilestone entry points alongside the shared generic ones rather than modifying them — init.phase-op and init.manager carry a CRITICAL blast radius (179 dependents, 24 processes) and stay byte-identical for their other callers. Admit flag:--fix, state:fallow-enabled and state:git-create-tag, each with a consuming section and a fact its own entry point computes. Both sections had the resolver-in-body hazard: the fallow config-gate and the git.create_tag check each sat inside the very block being gated, so gating would have disabled the resolver that decides the gate. Both are hoisted into init and the bodies now consume the resolved fact. Co-Authored-By: Claude Opus 5 (1M context) <noreply@anthropic.com> * test(#2994): retarget code-review and milestone drift guards, fix two red tests Retarget guards that asserted on content moved into steps/, proving non-vacuity by blanking each step file and confirming failure. Also fixes two genuinely red tests found while working, per the no-defer rule: - workflow-fragments' frozen-vocabulary lock was missing state:ui-phase-active, so commit 7ef7f8336 shipped red. Lint and build both passed over it, which is why neither is sufficient verification. - code-review's quick.md capability-hook assertion carried a stale delimiter after the 18ff35d20 extraction. Co-Authored-By: Claude Opus 5 (1M context) <noreply@anthropic.com> * chore(#2994): fragmentize autonomous.md and admit state:plan-strategy-converge Five sections share one atom, the pattern plan-phase already uses for flag:--research-phase. The atom folds --converge OR --cross-ai into a single boolean in cmdInitAutonomous so the grammar stays operator-free. cmdInitAutonomous is additive; init.milestone-op, init.manager and init.phase-op are untouched and still consumed. The $PLAN_STRATEGY bash resolver is deliberately retained — ungated local-planning bullets still read it, so the init-side fact supplements it rather than replacing it. converge-fail-fast required splitting one bash fence so the always-run CONVERGENCE_ARGS construction stays outside the marker. All three flag-absent fallbacks were left outside their markers. Co-Authored-By: Claude Opus 5 (1M context) <noreply@anthropic.com> * chore(#2994): fragmentize review and discuss-phase-assumptions Admit state:reviewer-instances-configured (two peripheral notes share it; the core reviewer-lane dispatch stays unmarked — it is the workflow's primary always-evaluated logic, not an optional branch) and state:auto-advance-active, which folds --auto OR two config keys into one boolean so the grammar stays operator-free. discuss-phase-assumptions was the highest-risk edit in this PR. Its auto_advance step is a full if/elif/else; gating it whole would have deleted the flag-absent fallback needed exactly when --auto is off. Split verified exact: resolvers 636-651 and the 'End here' fallback 668-669 both stay outside the marker; only 653-667 is gated. Adds emitted-drift acks for the two files that grew — review.md (+55 B) and autonomous.md (+737 B from 80799211c, which had none and would have red-gated the push. Co-Authored-By: Claude Opus 5 (1M context) <noreply@anthropic.com> * chore(#2994): fragmentize docs-update, update, transition and new-milestone Part A Completes the 13-workflow rollout. Three of these had no init call at all and gained a dedicated entry point plus their first gsd_run query line. Admits state:is-monorepo and adds state:next-channel, state:workstream-active and state:flat-mode. Vocabulary 26 -> 30 atoms. Part A of new-milestone applies when NO workstream is active — the negation of state:workstream-active. Rather than teach the grammar negation, which is the Greenspun drift the frozen list exists to prevent, it gets a separate positively-phrased atom whose fact is the inverse. Part B, which always runs, stays outside the marker. flag:--verify-only is deliberately NOT admitted: docs-update has no contiguous purely-additive region for it, and an atom without a consuming section is dead vocabulary. Evidence recorded in the slice report. update.md reuses its existing resolved $GSD_TOOLS rather than prepending the canonical preamble, which would have clobbered it. Co-Authored-By: Claude Opus 5 (1M context) <noreply@anthropic.com> * fix(#2994): stop automated-ui-verification re-resolving its own gate, retire dead vocabulary Two defects the new tests caught. The automated-ui-verification step re-ran gsd_run loop render-hooks and recomputed UI_PHASE_ACTIVE inside a body that is only read when that fact is already true — the circular self-disabling pattern this design forbids, introduced by 3c654b168. cmdInitVerifyWork now exposes ui_phase_active and the step consumes it. Its launcher preamble goes too: no gsd_run remains. The Playwright-MCP check stays as prose — that is live session state. Dead vocabulary predating this PR: flag:--full and state:needs-codebase-map were admitted with a gate-1 claim that never materialized. flag:--full is removed, redundant once quick folds it into discuss/research/validate. state:needs-codebase-map gets the real consumer it always lacked, gating new-project's codebase-map offer. Vocabulary 30 -> 29, and no atom is now without a consuming section. Co-Authored-By: Claude Opus 5 (1M context) <noreply@anthropic.com> * test(#2994): add the atom-admission, inversion and resolver-hoist gates The two existing parity guards prove vocabulary/predicate symmetry but never that a fact is computed — an atom no cmdInit* assembles evaluates false forever. These close that hole: - per-atom satisfiability for all 29 atoms, plus an anti-vacuity assertion so the loop cannot silently cover zero atoms - dead-vocabulary check against the shipped manifest - inversion guard: the flag-absent fallbacks in discuss-phase-assumptions and verify-work must stay outside their markers - data-driven resolver-hoist guard over the shipped manifest, so a future extraction cannot reintroduce the circular class - compound-fold coverage (--full, --cross-ai, --rc, config-only --auto) - null-vs-[] degraded/computed distinction, and flag value shapes Also repairs the frozen-vocabulary lock, which was stale and red for the seven atoms earlier commits on this branch shipped. Co-Authored-By: Claude Opus 5 (1M context) <noreply@anthropic.com> * docs(#2994): add changeset for the fragment-model rollout Co-Authored-By: Claude Opus 5 (1M context) <noreply@anthropic.com> * test(#2994): cite the issue on the two new allow-test-rule exemptions ADR-456 requires an issue ref on the same line as the annotation. Co-Authored-By: Claude Opus 5 (1M context) <noreply@anthropic.com> * docs(#2994): correct the atom-count claims after retiring flag:--full The vocabulary doc comments still said 30 entries; it is 29 since flag:--full was removed as dead vocabulary. Co-Authored-By: Claude Opus 5 (1M context) <noreply@anthropic.com> * fix(#2994): dedupe the phase-fallback block and harden --ws parsing Review findings. MAJOR: the three new init entry points each pasted a verbatim copy of the guardedFindPhase/guardedGetRoadmapPhase fallback, taking the repo from four copies to seven — DEFECT.GENERATIVE-FIX. Extracted applyRoadmapFallback and folded six of the seven; each call site keeps its own field-set via a closure. Duplication removed rather than papered over with a parity test. cmdInitPhaseOp stays out: its fallback omits has_reviews, so it is not a byte-identical copy, and it is CRITICAL-radius. LOW, pre-existing: GSD_WS captured [^[:space:]]+ and expands unquoted, so a workstream name holding glob metacharacters would expand against the filesystem. Narrowed to [A-Za-z0-9._-]+. The unquoted expansion is kept — it must word-split into two args and vanish when empty. Also restores the vocabulary ordering convention, and fixes a masked test bug the mandated run surfaced: the flag-forwarding guard checked only the first init line per workflow, but new-milestone has two, so a real failure was reporting exit 0. Co-Authored-By: Claude Opus 5 (1M context) <noreply@anthropic.com> * fix(#2994): drop the stale new-milestone emitted-drift ack new-milestone.md was acked for a +406 B growth measured against an intermediate commit. Net against origin/next it SHRANK by 8 bytes, so nothing needed the ack and it explained nothing — which the differential attribution check reports as a stale acknowledgment, not a pass. update.md's entry stays: it genuinely grew +703 B. Co-Authored-By: Claude Opus 5 (1M context) <noreply@anthropic.com> * fix(#2994): resolve the 15 failures from the full matrix run All 15 were real and identical on both lanes. REAL REGRESSION: autonomous.md hit 41479 chars against the #2196 guard's 40960 cap — a CHARS cap distinct from the LARGE tier byte cap, which the five section stubs pushed it over. Extracted the 3a.5 UI Design Contract body to references/; now 39968 chars, and the file nets -795 B vs base, so its growth ack is deleted rather than left stale. REAL DEFECT: docs referenced /gsd-transition, which is not a live registered command. Reworded. STALE FIXTURE: the emission byte-identity test hardcoded two marked workflows; this branch legitimately marks fifteen. Fixture corrected — the source was right. The rest were drift guards over the eight workflows the earlier sweep did not cover, retargeted at where the content now lives with non-vacuity proven by blanking each step file and confirming failure. The GSD_WS forwarding guard was checked as a possible real break and is not one: the charclass narrowing is intact and forwarding works end to end. Co-Authored-By: Claude Opus 5 (1M context) <noreply@anthropic.com> * fix(#2994): drop the ack for a newly-added reference file A new file's emitted ripple is attributable to the diff that adds it, so the acknowledgment explained nothing and the differential check reports it as stale. Removing the last entry removes the fragment — an empty one signals nothing. Co-Authored-By: Claude Opus 5 (1M context) <noreply@anthropic.com> * fix(#2994): retarget the UI-contract guards and clear two transitive advisories The §3a.5 extraction that brought autonomous.md under the #2196 char cap moved its body to references/autonomous-ui-design-contract.md, so ten guards in autonomous-ui-steps and check-ui-safety-gate were asserting it against the host. Retargeted via a combined read, each proven non-vacuous by blanking the reference file and confirming failure. This class had already bitten twice on this branch because each sweep was scoped to the workflows touched at that moment, so this one was exhaustive: ~70 test files across all 13 workflows, zero further broken or vacuous assertions found. Also clears two high transitive advisories the matrix flagged on one lane — fast-uri GHSA-7p8r-x3mc-p8w7 and three ip-address SSRF/trust-boundary issues. Both pre-date this branch: package-lock.json was untouched until now, so the production tree was byte-identical to the base. Lockfile-only, package.json unchanged, verified against a real npm ci install. Co-Authored-By: Claude Opus 5 (1M context) <noreply@anthropic.com> * chore(#2994): backfill changeset pr number to 3030 --------- Co-authored-by: sim <sim@local> Co-authored-by: Claude Opus 5 (1M context) <noreply@anthropic.com>
346 lines
15 KiB
JavaScript
346 lines
15 KiB
JavaScript
/**
|
|
* #2287 — deferred-items.md has no reader anywhere in gsd-core.
|
|
*
|
|
* The SCOPE BOUNDARY convention (`agents/gsd-executor.md`) instructs the
|
|
* executor to log out-of-scope discoveries to `deferred-items.md` inside the
|
|
* phase directory. Nothing read that file back: `cmdAuditUat` (src/uat.cts)
|
|
* filtered phase-directory files down to `*-UAT.md` / `*-VERIFICATION.md`
|
|
* only, and the `forensic_audit` workflow step (gsd-core/workflows/
|
|
* progress.md) ran 6 checks, none of which globbed the phase-directory
|
|
* `deferred-items.md` path. An entry written there was permanently invisible.
|
|
*
|
|
* This fix:
|
|
* - `cmdAuditUat` gains a `deferred-items.md` scan per phase directory,
|
|
* surfacing every UNRESOLVED entry as a `type: 'deferred'` result. An
|
|
* entry is resolved only when it carries an explicit `status: resolved`
|
|
* field (mirroring the established `## Gaps` convention from #2286) — a
|
|
* missing/garbled status fails safe and is surfaced.
|
|
* - `forensic_audit` gains a 7th check that globs the same path and reports
|
|
* unresolved entries with the same ✓/⚠ semantics as the other 6 checks.
|
|
*
|
|
* `deferred-items.md` remains the single source of truth — no duplicate
|
|
* `.planning/todos/pending/*.md` entry is required.
|
|
*/
|
|
|
|
'use strict';
|
|
|
|
const { test, describe, beforeEach, afterEach } = require('node:test');
|
|
const assert = require('node:assert/strict');
|
|
const fs = require('node:fs');
|
|
const path = require('node:path');
|
|
const fc = require('./helpers/fast-check-setup.cjs');
|
|
const { runGsdTools, createTempProject, cleanup } = require('./helpers.cjs');
|
|
const { parseDeferredItems } = require('../gsd-core/bin/lib/uat.cjs');
|
|
|
|
// ─── cmdAuditUat behavioral coverage ───────────────────────────────────────
|
|
|
|
describe('#2287 cmdAuditUat: deferred-items.md awareness', () => {
|
|
let tmpDir;
|
|
|
|
beforeEach(() => {
|
|
tmpDir = createTempProject();
|
|
});
|
|
|
|
afterEach(() => {
|
|
cleanup(tmpDir);
|
|
});
|
|
|
|
test('no deferred-items.md present (0 entries) → no results, no false positive', () => {
|
|
const phaseDir = path.join(tmpDir, '.planning', 'phases', '01-foundation');
|
|
fs.mkdirSync(phaseDir, { recursive: true });
|
|
fs.writeFileSync(path.join(phaseDir, '.gitkeep'), '');
|
|
|
|
const result = runGsdTools('audit-uat --raw', tmpDir);
|
|
assert.ok(result.success, `Command failed: ${result.error}`);
|
|
|
|
const output = JSON.parse(result.output);
|
|
assert.deepStrictEqual(output.results, []);
|
|
assert.strictEqual(output.summary.total_items, 0);
|
|
assert.strictEqual(output.summary.total_files, 0);
|
|
});
|
|
|
|
test('deferred-items.md with only a resolved entry (0 unresolved) → no result surfaced', () => {
|
|
const phaseDir = path.join(tmpDir, '.planning', 'phases', '01-foundation');
|
|
fs.mkdirSync(phaseDir, { recursive: true });
|
|
|
|
fs.writeFileSync(path.join(phaseDir, 'deferred-items.md'), [
|
|
'## Deferred Items',
|
|
'',
|
|
'- Already handled unrelated lint warning.',
|
|
' status: resolved',
|
|
].join('\n'));
|
|
|
|
const result = runGsdTools('audit-uat --raw', tmpDir);
|
|
assert.ok(result.success, `Command failed: ${result.error}`);
|
|
|
|
const output = JSON.parse(result.output);
|
|
assert.deepStrictEqual(output.results, [],
|
|
'a fully-resolved deferred-items.md must not surface any result');
|
|
assert.strictEqual(output.summary.total_items, 0);
|
|
});
|
|
|
|
test('deferred-items.md with 1 unresolved entry → surfaced in structured JSON output', () => {
|
|
const phaseDir = path.join(tmpDir, '.planning', 'phases', '01-foundation');
|
|
fs.mkdirSync(phaseDir, { recursive: true });
|
|
|
|
fs.writeFileSync(path.join(phaseDir, 'deferred-items.md'), [
|
|
'## Deferred Items',
|
|
'',
|
|
'- Found an unrelated pre-existing test failure in `some-other-module` while working on',
|
|
' this phase\'s task. Out of scope for this task — logged here per SCOPE BOUNDARY.',
|
|
].join('\n'));
|
|
|
|
const result = runGsdTools('audit-uat --raw', tmpDir);
|
|
assert.ok(result.success, `Command failed: ${result.error}`);
|
|
|
|
const output = JSON.parse(result.output);
|
|
assert.strictEqual(output.summary.total_items, 1);
|
|
assert.strictEqual(output.summary.total_files, 1);
|
|
assert.strictEqual(output.summary.by_category.deferred, 1);
|
|
assert.strictEqual(output.summary.by_phase['01'], 1);
|
|
|
|
const deferredResult = output.results.find(r => r.type === 'deferred');
|
|
assert.ok(deferredResult, 'a deferred-typed result must be present');
|
|
assert.strictEqual(deferredResult.phase, '01');
|
|
assert.strictEqual(deferredResult.file, 'deferred-items.md');
|
|
assert.strictEqual(
|
|
deferredResult.file_path,
|
|
'.planning/phases/01-foundation/deferred-items.md',
|
|
);
|
|
assert.strictEqual(deferredResult.items.length, 1);
|
|
assert.match(deferredResult.items[0].name, /unrelated pre-existing test failure/);
|
|
assert.strictEqual(deferredResult.items[0].result, 'unresolved');
|
|
assert.strictEqual(deferredResult.items[0].category, 'deferred');
|
|
});
|
|
|
|
test('deferred-items.md with 2+ entries (mixed resolved/unresolved) → only unresolved surfaced', () => {
|
|
const phaseDir = path.join(tmpDir, '.planning', 'phases', '01-foundation');
|
|
fs.mkdirSync(phaseDir, { recursive: true });
|
|
|
|
fs.writeFileSync(path.join(phaseDir, 'deferred-items.md'), [
|
|
'## Deferred Items',
|
|
'',
|
|
'- First unrelated finding, still open.',
|
|
'- Second unrelated finding, also still open.',
|
|
'- Third finding, already fixed separately.',
|
|
' status: resolved',
|
|
].join('\n'));
|
|
|
|
const result = runGsdTools('audit-uat --raw', tmpDir);
|
|
assert.ok(result.success, `Command failed: ${result.error}`);
|
|
|
|
const output = JSON.parse(result.output);
|
|
const deferredResult = output.results.find(r => r.type === 'deferred');
|
|
assert.ok(deferredResult);
|
|
assert.strictEqual(deferredResult.items.length, 2,
|
|
'exactly the 2 unresolved entries must surface; the resolved 3rd must not');
|
|
const names = deferredResult.items.map(i => i.name);
|
|
assert.ok(names.some(n => n.includes('First unrelated finding')));
|
|
assert.ok(names.some(n => n.includes('Second unrelated finding')));
|
|
assert.ok(!names.some(n => n.includes('Third finding')));
|
|
});
|
|
|
|
test('deferred entries surface across multiple phase directories', () => {
|
|
const phase1 = path.join(tmpDir, '.planning', 'phases', '01-foundation');
|
|
const phase2 = path.join(tmpDir, '.planning', 'phases', '02-auth');
|
|
fs.mkdirSync(phase1, { recursive: true });
|
|
fs.mkdirSync(phase2, { recursive: true });
|
|
|
|
fs.writeFileSync(path.join(phase1, 'deferred-items.md'), [
|
|
'## Deferred Items',
|
|
'',
|
|
'- Phase 1 unrelated finding.',
|
|
].join('\n'));
|
|
fs.writeFileSync(path.join(phase2, 'deferred-items.md'), [
|
|
'## Deferred Items',
|
|
'',
|
|
'- Phase 2 unrelated finding.',
|
|
].join('\n'));
|
|
|
|
const result = runGsdTools('audit-uat --raw', tmpDir);
|
|
assert.ok(result.success, `Command failed: ${result.error}`);
|
|
|
|
const output = JSON.parse(result.output);
|
|
const deferredResults = output.results.filter(r => r.type === 'deferred');
|
|
assert.strictEqual(deferredResults.length, 2);
|
|
assert.strictEqual(output.summary.total_items, 2);
|
|
assert.strictEqual(output.summary.by_phase['01'], 1);
|
|
assert.strictEqual(output.summary.by_phase['02'], 1);
|
|
});
|
|
|
|
test('an entry with a garbled/missing status fails safe and is surfaced (not silently dropped)', () => {
|
|
const phaseDir = path.join(tmpDir, '.planning', 'phases', '01-foundation');
|
|
fs.mkdirSync(phaseDir, { recursive: true });
|
|
|
|
fs.writeFileSync(path.join(phaseDir, 'deferred-items.md'), [
|
|
'## Deferred Items',
|
|
'',
|
|
'- An entry with no status field at all.',
|
|
].join('\n'));
|
|
|
|
const result = runGsdTools('audit-uat --raw', tmpDir);
|
|
assert.ok(result.success, `Command failed: ${result.error}`);
|
|
|
|
const output = JSON.parse(result.output);
|
|
assert.strictEqual(output.summary.total_items, 1,
|
|
'missing status must SURFACE the entry, not silently drop it');
|
|
});
|
|
|
|
test('existing UAT/VERIFICATION scanning is unchanged when a deferred-items.md is also present', () => {
|
|
const phaseDir = path.join(tmpDir, '.planning', 'phases', '01-foundation');
|
|
fs.mkdirSync(phaseDir, { recursive: true });
|
|
|
|
fs.writeFileSync(path.join(phaseDir, '01-UAT.md'), [
|
|
'---',
|
|
'status: testing',
|
|
'phase: 01-foundation',
|
|
'started: 2025-01-01T00:00:00Z',
|
|
'updated: 2025-01-01T00:00:00Z',
|
|
'---',
|
|
'',
|
|
'## Tests',
|
|
'',
|
|
'### 1. Login Form',
|
|
'expected: Form displays with email and password fields',
|
|
'result: pending',
|
|
].join('\n'));
|
|
|
|
fs.writeFileSync(path.join(phaseDir, 'deferred-items.md'), [
|
|
'## Deferred Items',
|
|
'',
|
|
'- An unrelated out-of-scope finding.',
|
|
].join('\n'));
|
|
|
|
const result = runGsdTools('audit-uat --raw', tmpDir);
|
|
assert.ok(result.success, `Command failed: ${result.error}`);
|
|
|
|
const output = JSON.parse(result.output);
|
|
assert.strictEqual(output.results.length, 2, 'both the UAT file and deferred-items.md must surface as separate results');
|
|
const uatResult = output.results.find(r => r.type === 'uat');
|
|
const deferredResult = output.results.find(r => r.type === 'deferred');
|
|
assert.ok(uatResult, 'existing uat-type result must still be present');
|
|
assert.strictEqual(uatResult.items.length, 1);
|
|
assert.strictEqual(uatResult.items[0].result, 'pending');
|
|
assert.ok(deferredResult, 'new deferred-type result must be present');
|
|
assert.strictEqual(deferredResult.items.length, 1);
|
|
});
|
|
});
|
|
|
|
// ─── forensic_audit workflow-prose source-contract guard ──────────────────
|
|
|
|
// #2994 fragmentization moved the --forensic-gated forensic_audit step out of
|
|
// progress.md into gsd-core/workflows/progress/steps/forensic-audit.md behind
|
|
// a section marker. Read that step file directly — it is the sole remaining
|
|
// source of the forensic_audit step body these guards assert on.
|
|
const PROGRESS_MD = path.join(__dirname, '..', 'gsd-core', 'workflows', 'progress', 'steps', 'forensic-audit.md');
|
|
|
|
describe('#2287 progress.md forensic_audit: deferred-items.md contract', () => {
|
|
const content = fs.readFileSync(PROGRESS_MD, 'utf-8');
|
|
const stepStart = content.indexOf('<step name="forensic_audit">');
|
|
const stepEnd = content.indexOf('</step>', stepStart);
|
|
const section = stepStart !== -1 && stepEnd !== -1 ? content.slice(stepStart, stepEnd) : '';
|
|
|
|
test('forensic_audit step exists', () => {
|
|
assert.notEqual(stepStart, -1, 'progress.md (or its extracted progress/steps/forensic-audit.md) must contain the forensic_audit step');
|
|
});
|
|
|
|
test('forensic_audit now runs 7 checks (was 6) and globs deferred-items.md', () => {
|
|
assert.ok(/running 7 deep checks/i.test(section),
|
|
'forensic_audit must advertise 7 deep checks (was 6) now that deferred-items.md is read');
|
|
assert.ok(/\.planning\/phases\/\*\/deferred-items\.md/.test(section),
|
|
'forensic_audit must glob .planning/phases/*/deferred-items.md');
|
|
});
|
|
|
|
test('the new check reports unresolved deferred items with the same ✓/⚠ semantics as the other checks', () => {
|
|
assert.ok(/check\s*7/i.test(section),
|
|
'a 7th check must be present');
|
|
assert.ok(/unresolved deferred items/i.test(section),
|
|
'the check must be framed around unresolved deferred items');
|
|
assert.ok(/✓[^\n]*no unresolved deferred items/i.test(section),
|
|
'the check must emit a ✓ pass line when no unresolved deferred items exist');
|
|
assert.ok(/⚠[^\n]*unresolved deferred items found/i.test(section),
|
|
'the check must emit a ⚠ warning line when unresolved deferred items exist');
|
|
});
|
|
|
|
test('an entry is resolved only via an explicit status: resolved field (fail-safe otherwise)', () => {
|
|
assert.ok(/status:\s*resolved/i.test(section),
|
|
'the resolved/unresolved parsing rule must be documented in the step prose');
|
|
});
|
|
|
|
test('the verdict summary now gates on 7 checks (was 6)', () => {
|
|
assert.ok(/after all 7 checks/i.test(section),
|
|
'the verdict section must say "after all 7 checks"');
|
|
assert.ok(/if all 7 checks passed/i.test(section),
|
|
'the verdict section must say "if all 7 checks passed"');
|
|
assert.ok(!/after all 6 checks/i.test(section) && !/if all 6 checks passed/i.test(section),
|
|
'stale "6 checks" phrasing must not remain in the step');
|
|
});
|
|
});
|
|
|
|
// ─── parseDeferredItems property test ──────────────────────────────────────
|
|
|
|
describe('#2287 parseDeferredItems: property (status: resolved fail-safe)', () => {
|
|
// Single-line entry text: no newlines (would break bullet-entry splitting),
|
|
// non-empty after trim, and never itself SHAPED like a `status:` field line
|
|
// (that would be indistinguishable from a real field regardless of intent).
|
|
const plainText = fc.string({ minLength: 1, maxLength: 40 })
|
|
.map((s) => s.replace(/[\r\n]/g, ' ').trim())
|
|
.filter((s) => s.length > 0 && !/^status:/i.test(s));
|
|
|
|
// Decoy: entry text that CONTAINS a `status: resolved`-shaped substring
|
|
// mid-line (not at line start) — must never be misread as a resolved
|
|
// marker, since extractGapEntryFields only recognises a field anchored to
|
|
// the START of its own trimmed line (see parseDeferredItems' doc comment).
|
|
const decoyText = plainText.map((s) => `${s} status: resolved trailing note`);
|
|
|
|
const textArb = fc.oneof(plainText, decoyText);
|
|
const entryArb = fc.record({ text: textArb, resolved: fc.boolean() });
|
|
|
|
test('property: an entry is surfaced iff it is NOT marked status: resolved; surfaced count == non-resolved count', () => {
|
|
fc.assert(
|
|
fc.property(
|
|
fc.array(entryArb, { maxLength: 20 }),
|
|
(rawEntries) => {
|
|
// Index-prefix for uniqueness so surfaced items can be mapped back
|
|
// to their source entry unambiguously even with colliding random text.
|
|
const entries = rawEntries.map((e, i) => ({ text: `E${i}_${e.text}`, resolved: e.resolved }));
|
|
|
|
const lines = ['## Deferred Items', ''];
|
|
for (const e of entries) {
|
|
lines.push(`- ${e.text}`);
|
|
if (e.resolved) lines.push(' status: resolved');
|
|
}
|
|
const content = lines.join('\n');
|
|
|
|
const items = parseDeferredItems(content);
|
|
const surfacedNames = new Set(items.map((it) => it.name));
|
|
|
|
const expectedUnresolved = entries.filter((e) => !e.resolved);
|
|
const expectedResolved = entries.filter((e) => e.resolved);
|
|
|
|
// Total surfaced count equals the count of non-resolved entries.
|
|
assert.strictEqual(items.length, expectedUnresolved.length);
|
|
|
|
// Every non-resolved entry IS surfaced (including status:-shaped
|
|
// decoy substrings embedded mid-line — those must not flip the
|
|
// outcome).
|
|
for (const e of expectedUnresolved) {
|
|
assert.ok(surfacedNames.has(e.text), `expected unresolved entry to surface: ${e.text}`);
|
|
}
|
|
|
|
// No status:-resolved entry is EVER surfaced.
|
|
for (const e of expectedResolved) {
|
|
assert.ok(!surfacedNames.has(e.text), `status: resolved entry must never surface: ${e.text}`);
|
|
}
|
|
|
|
// Every returned item carries the fixed deferred category/result shape.
|
|
for (const item of items) {
|
|
assert.strictEqual(item.result, 'unresolved');
|
|
assert.strictEqual(item.category, 'deferred');
|
|
}
|
|
}
|
|
)
|
|
);
|
|
});
|
|
});
|