* test(#1168): make phase-6 gate un-gameable — reject empty stubs + require loop shrink The migration assertion previously checked only role==feature, so a registration-only stub (empty hooks, logic left inline) would turn the gate green while phase 6 stayed incomplete — the exact false-completion pattern this gate exists to prevent. Strengthen it: each ADR-named feature must OWN its behavior (>=1 hook, or a command family); and plan-phase.md/execute-phase.md must shrink strictly below their frozen pre-phase-6 sizes (94519/93166 LF bytes), which also defeats double-run gaming (declare a hook but keep the inline block -> file does not shrink -> red). Gate now 5 pass / 4 fail (orphaned execute:wave:post, empty/unregistered features, config-key leaks, no shrink). Green is now reachable only by REAL migration. Refs #1168, #1169. Co-Authored-By: Claude Opus 4.8 <noreply@anthropic.com> * feat(#1169): migrate gap-analysis to a Capability (plan:post gate) First real ADR-857 phase-6 migration (pattern-defining tracer). gap-analysis moves from an inline post_planning_gaps branch in plan-phase.md to a real plan:post gate Capability: - capabilities/gap-analysis/capability.json: role:feature, plan:post gate (when=workflow.post_planning_gaps, blocking:false advisory), OWNS workflow.post_planning_gaps (federated out of central schema). - plan-phase.md: inline config-get + gsd_run gap-analysis block replaced with a plan:post render-hooks call site dispatching the gate; file shrinks 94519->93279. - src/check-command-router.cts: cmdGapAnalysisPlanPost runs the real gap analysis via gap-checker. - post_planning_gaps removed from central manifest; resolves via federated config (default true preserved). - tests/post-planning-gaps-2493: re-pointed to assert capability ownership. Verified: gate 5 pass / 4 fail (gap-analysis cleared from migration, plan:post-orphan, config-leak, and plan-phase shrink checks); loadConfig still returns post_planning_gaps=true; check command runs real analysis; 392/392 in the config/registry/federation/router net. Refs #1169. Co-Authored-By: Claude Opus 4.8 <noreply@anthropic.com> * feat(#1169): migrate profile-pipeline to a command-family Capability ADR-857 Decision 7: profile-pipeline becomes a command-family Capability (like audit/intel/graphify). capabilities/profile-pipeline/capability.json declares an 8-command family (scan-sessions, extract-messages, profile-sample, write-profile, profile-questionnaire, generate-dev-preferences, generate-claude-profile, generate-claude-md) backed by a new gsd-core/bin/lib/profile-pipeline-command-router.cjs; the inline case arms are removed from gsd-tools.cjs. Owns profile-pipeline.enabled (federated). Verified: registry shows role:feature with commands.length=8; scan-sessions/profile-sample run live via the family; gate cleared profile-pipeline from the empty-stub failure (only tdd/schema-gate/drift remain); 296/296 registry+inventory+gsd-tools tests; lint 0 errors. Refs #1169. Co-Authored-By: Claude Opus 4.8 <noreply@anthropic.com> * fix(#1167): wire execute:wave:post + implement ui.safety-gate check Revives the second dead gate from #1167: ui.gates@execute:wave:post was declared but never dispatched AND its check.query (ui.safety-gate) was unimplemented. Adds the per-wave execute:wave:post render-hooks call site in execute-phase.md (fires after each wave's merge/cleanup, before the next forks) and implements cmdUiSafetyGate (frontend + UI-SPEC aware, mirrors cmdUiPlanGate) in check-command-router. +17 regression tests. Verified: phase-6 orphaned-points conformance test now PASSES (gate 6 pass / 3 fail); ui-safety-gate routable in dot+hyphen forms; check-ui-safety-gate 17/17, check-ui-plan-gate 18/18; lint 0 errors. Refs #1167, #1168. Co-Authored-By: Claude Opus 4.8 <noreply@anthropic.com> * feat(#1169): migrate drift (schema + codebase) to execute:wave:post gates Removes the inline schema_drift_gate + codebase_drift_gate steps (77 lines) from execute-phase.md; drift becomes a Capability with two execute:wave:post gates (verify.schema-drift blocking, verify.codebase-drift advisory) dispatched via the per-wave render-hooks call site. check-command-router routes verify.schema-drift / verify.codebase-drift to the real detectors. Federates workflow.drift_threshold / drift_action / schema_drift_gate out of central. Also fixes the execute:wave:post dispatch prose to run NON-blocking (advisory) gates too — the prior version only ran blocking gates, which would have silently dropped the codebase-drift advisory after its inline step was removed. Behavior preserved. Verified: gate 7 pass / 2 fail (drift cleared from stub + config-leak; execute-phase.md 92297 < 93166 frozen -> shrink passes); both drift checks run real detection; loadConfig defaults preserved (threshold=3, action=warn, gate=true); drift-detection 56/56 + schema-drift 34/34; lint 0 errors. Refs #1169. Co-Authored-By: Claude Opus 4.8 <noreply@anthropic.com> * feat(#1169): migrate tdd to a Capability (plan:pre contribution + execute:post gate) tdd becomes a real Capability: a plan:pre contribution injects the <tdd_mode_active> planner guidance (rendered from PLAN_PRE_HOOKS_JSON like security's contribution), and an execute:post gate (tdd.review-checkpoint, advisory) runs the real end-of-phase RED/GREEN review via a new check-command handler. Inline tdd_mode reads + the inline planner block + the tdd_review_checkpoint step are removed; workflow.tdd_mode is federated out of central. The MVP+TDD per-task RED-commit gate is preserved — TDD_MODE is now derived from the execute:post hooks (capId==tdd active), not an inline config-get. BEHAVIOR CHANGE (documented, not silent): the --tdd CLI flag now persists workflow.tdd_mode=true via config-set instead of being per-invocation. Rationale: tdd is now a config-toggled Capability, and env vars do not persist across the workflow's separate bash blocks (config does), so an ephemeral override isn't cleanly achievable; --tdd therefore enables the tdd capability, consistent with how all capabilities are toggled. Verified: gate 7 pass / 2 fail (tdd cleared from stub + config-leak; plan-phase + execute-phase both < frozen sizes); contribution injection + execute:post gate dispatch wired; MVP+TDD gate preserved; tdd.review-checkpoint runs real review; full unit suite 556/0; lint 0 errors. Refs #1169. Co-Authored-By: Claude Opus 4.8 <noreply@anthropic.com> * feat(#1169): migrate schema-gate to a plan:pre contribution Capability The plan-time schema-push detection (former plan-phase.md §5.7) becomes a schema-gate Capability: a plan:pre contribution (into:planner, when:workflow.schema_push_detection) whose fragment carries the full ORM-detection + [BLOCKING] schema-push-task injection logic, rendered into the planner via the existing plan:pre render-hooks dispatch. The inline §5.7 block is removed (plan-phase.md 94519->90445). workflow.schema_push_detection is a new capability-owned (federated) key, default true. (The execute-side schema-drift gate was migrated separately into the drift capability.) Verified: registry inlines the fragment (len 2704) so it is actually delivered at plan:pre; gate 8 pass / 1 fail — all 5 ADR-named features now real Capabilities, only the config-leak test remains (intel/security, next unit). Refs #1169. Co-Authored-By: Claude Opus 4.8 <noreply@anthropic.com> * feat(#1169): close the 3 capability config-key leaks — phase-6 gate now GREEN Removes the last inline config-get reads of capability-owned keys from plan-phase.md. security_asvs_level/security_block_on now flow through the security plan:pre contribution via a new loop-resolver configValues mechanism (resolves declared config keys with the same 4-level precedence as activation and attaches them to the rendered hook); the §5.55 banner reads them from PLAN_PRE_HOOKS_JSON. intel.enabled becomes a real intel plan:pre step (ref.command: intel api-surface) dispatched via render-hooks; the inline intel branch is gone. gen-capability-registry now validates ref.command as a third dispatch shape. Verified: phase-6 capstone conformance gate is FULLY GREEN (9/0); 3 leaks gone (grep=0); security configValues resolve to {2,medium}/default {1,high}; intel step present only when enabled; loop-render-hooks 62/0, capability-registry 287/0, capability-state/federated-config 113/0; lint 0 errors. Closes the migration half of #1169. Refs #1139, #1167, #1168. Co-Authored-By: Claude Opus 4.8 <noreply@anthropic.com> * fix(#1169): address adversarial review — restore schema-drift block, generic planner injection, uniform gate contract Adversarial review caught 2 real regressions the green gate missed: (1) schema-drift no longer blocked — the execute:wave:post dispatch read GATE_RESULT.block but verify.schema-drift emitted drift_detected/blocking, and onError:skip wrongly bypassed positive blocks; (2) only tdd's plan:pre contribution was injected into the planner, dropping schema-gate's schema-push detection and security's threat-model guidance. Fixes: (A) every gate check returns a uniform boolean 'block' under --raw (the dispatch form), with advisory gates (tdd/gap) carrying their report in 'message'; (B) gate-dispatch contract corrected at all sites — onError governs command errors only, a blocking gate's positive block always halts; (C) generic planner injection of all plan:pre contributions where into=='planner' (tdd + schema-gate + security incl configValues); (D) two new conformance assertions: planner contributions injected generically + every gate check.query returns boolean block under --raw. Verified: gate 11/11; all 6 gate checks return boolean block under --raw; full suite 595/0; lint 0 errors. Refs #1167, #1168, #1169. Co-Authored-By: Claude Opus 4.8 <noreply@anthropic.com> * fix(#1169): restore MVP+TDD end-of-phase blocking escalation (2nd adversarial pass) The migrated tdd execute:post gate is statically blocking:false, but the contract (references/execute-mvp-tdd.md + CONTEXT.md) requires the end-of-phase TDD review to ESCALATE from advisory to blocking when MVP_MODE && TDD_MODE && a TDD plan misses a RED/GREEN commit. The migration prose had downgraded this to a 'strong advisory recommendation' — silent loss of the blocking escalation. Restore it: the tdd-gate dispatch now refuses to mark the phase complete (Phase blocked message) under MVP+TDD when GATE_RESULT.block is true; advisory otherwise. Also strengthen tests/execute-mvp-tdd-gate.test.cjs: hasBlockingEscalation previously matched any line with 'blocking'+'mvp+tdd' (so 'advisory (blocking: false) ... under MVP+TDD' was a false green); now it requires the real refusal semantics ('refuse to mark the phase complete' / 'phase blocked'). Caught by 2nd adversarial review pass. Verified: execute-phase.md 92702 < 93166 frozen; mvp-tdd-gate + phase-6 gate 19/0; full suite green; lint 0 errors. Refs #1169. Co-Authored-By: Claude Opus 4.8 <noreply@anthropic.com> * fix(#1169): restore MVP+TDD proceed-block, codebase auto-remap, schema skip-flag (3rd adversarial pass) 3rd adversarial pass found 4 more silent regressions: (1) the tdd MVP+TDD 'refuse to mark complete' was nullified by a downstream 'ALWAYS proceed regardless of gate results' line — proceed is now conditional (stops on an active MVP+TDD block); (2) the test now asserts the proceed is NOT an unconditional override; (3) codebase-drift auto-remap (spawn gsd-codebase-mapper when drift_action=auto-remap) was dropped — the execute:wave:post advisory dispatch now consumes spawn_mapper/directive; (4) GSD_SKIP_SCHEMA_CHECK bypass was lost from the gate path — cmdVerifySchemaDrift now honors the env var (block:false when set). Verified: no unconditional proceed; GSD_SKIP_SCHEMA_CHECK=true -> block:false; gate 11/11 + mvp-tdd 9/9; full suite 569/0; lint 0; execute-phase.md 93109 < 93166. Refs #1169. Co-Authored-By: Claude Opus 4.8 <noreply@anthropic.com> * fix(#1169): init.cts reads federated config keys from nested path (4th adversarial pass) Config federation moved tdd_mode/research/nyquist_validation from flat config.<key> to nested config.workflow.<key>, but src/init.cts still read them flat — so init.plan-phase/init.execute-phase emitted tdd_mode:false / research_enabled:undefined / nyquist:undefined regardless of config (a public command-contract regression; the migrated loops use render-hooks so enforcement was unaffected). Read via config.workflow (type-safe Record cast). Now init reflects the same resolved values + federated defaults (research/nyquist default true) as the render-hooks path. Verified: build clean; init.plan-phase emits tdd_mode:true/research:false/nyquist:false for set config, defaults true for empty; full suite 591/0; lint 0. Refs #1169. Co-Authored-By: Claude Opus 4.8 <noreply@anthropic.com> * docs(#1169): add changeset for ADR-857 phase-6 completion (PR #1183) Co-Authored-By: Claude Opus 4.8 <noreply@anthropic.com> * fix(#1169): complete phase-6 migration fallout — restore TEXT_MODE, fix registry .claude leak, re-point stale workflow-contract tests The capability migration left real regressions and stale consumer tests that the per-module unit suite missed but the full cross-platform suite caught (27 failing tests): Real source regressions (fixed): - execute-phase.md lost its AskUserQuestion TEXT_MODE plain-text fallback when the inline schema_drift_gate step was removed — non-Claude runtimes would stall. Restored, and the execute:post gate-dispatch prose de-duplicated to cite the execute:wave:post contract (loop body shrinks below the frozen pre-phase-6 ceiling while keeping every onError/blocking nuance). - capabilities/tdd inline fragment hardcoded `@~/.claude/gsd-core/references/tdd.md`, baked verbatim into the committed capability-registry.cjs and leaked the install path on 11 non-Claude runtimes (registry .cjs is copied, not path-converted). Made the fragment path-free; regenerated the registry. The phase-6 conformance gate now guards this (no ~/.claude install path in any capability source or the generated registry). - plan-phase.md: removed a §5.7 stub re-added in error and routed Branch 2 to step 6 (schema-gate is a plan:pre capability, §5.7 is gone). Stale workflow-contract tests re-pointed to the capability dispatch they now must assert (behavior verified preserved in source first, assertions kept equal-or-stronger): bug-621 + bug-2851 (gap-analysis via gsd_run render-hooks plan:post + registry binding), feat-2527 (tdd_mode federated out of central), phase6-planning + plan-phase-ui-redirect (§5.6 bounded by ## 6.), plan-phase-drift-guard (intel when:intel.enabled skip branch). profile-pipeline-command-router.cjs un-ignored from eslint (hand-written, no TS source) + stale disable comments removed. Size baseline regenerated. Verified: full suite 15140 tests / 0 fail; lint 0 errors; conformance gate green legitimately. Refs #1139, #1167, #1168, #1169. Co-Authored-By: Claude Opus 4.8 <noreply@anthropic.com> * test(#1169): add ADR-857 E2E content-test coverage for the 12 loop points + capability deliverables Grounds the capability engine in behavioral E2E tests (drive the real render-hooks/check CLI + the real registry, assert typed result content — no source-grep), structured around what ADR-857 says to deliver. 207 tests; each genuineness-checked (flip the expectation, confirm it fails). Per-loop-point dispatch (7 files): empty-point negative-space across the 6 no-hook points; verify:post 3-step resolution+ordering+onError; plan:pre contribution/configValues + ui.plan-gate + intel; plan:post gap-analysis; execute:wave:post drift+ui gates via the check route (schema-drift block/skip, codebase-drift threshold BVA, auto-remap); execute:post tdd.review-checkpoint RED/GREEN; ship:pre security gate resolution + frontmatter-get predicate pieces. ADR-deliverable coverage (4 files): predicate boundary held (edge/prohibition probes stay core, not off-by-default Feature Capabilities — phase-6 exception); core loop runs with zero capabilities (all 12 points empty, init bundles resolve); contribution merge (multiple ordered <contribution from=> blocks); federated-config key removal on uninstall. federated-config allowlisted for its 3-file split (unit + integration + lifecycle). Refs #1139, #1167, #1168, #1169. Co-Authored-By: Claude Opus 4.8 <noreply@anthropic.com> * fix(#1169): remove dead drifted converter dups + address adversarial review Lint cleanup (root-caused, not waved off): src/runtime-artifact-conversion.cts carried 11 agent-converter functions (+5 orphaned consts/helpers) that were never exported, never called, and had silently DRIFTED from the live hand-authored copies in bin/install.js (one even referenced an undefined `claudeToCopilotTools`). Deleted the dead duplicates; install.js's live copies are untouched (it never imported these). Lint now 0 errors / 0 warnings. Adversarial-review (Codex) findings fixed: - HIGH: execute-phase.md TDD_MODE used `jq ... || echo false`, silently disabling the MVP+TDD blocking gate on jq-less runtimes. Reverted to the `node -e` form (node is guaranteed; matches the file's other node-e usages) so a missing optional tool can no longer fail-open a blocking safety path. - MEDIUM: federated-config-key-removal orphan-key test was vacuous (it skipped the orphan assertion). Now asserts the removed capability's key is genuinely not surfaced/validated after uninstall. - LOW: phase-6 conformance leak regex broadened to catch absolute-home and Windows-backslash `.claude/(gsd-core|commands|agents|hooks)` paths, not only `~`/`$HOME` forward-slash forms. - LOW: bug-2851 plan:post dispatch assertion now requires `--raw` (matched its stated contract). - nit: plan-pre intel-step test duplicate assertion replaced with a distinct structured-output check. Size baseline regenerated (execute-phase.md 93089 < 93166 frozen). Refs #1167, #1169. Co-Authored-By: Claude Opus 4.8 <noreply@anthropic.com> * test(#1169): make runtime-homes-descriptor-drive titles environment-independent The descriptor-equivalence test embedded the absolute golden config path (`os.homedir()`-derived) directly in each `test(...)` title, so titles differed between macOS (`/Users/x/.claude`) and Docker (`/home/gsdtest/.claude`). Every test PASSES on both platforms (15885/0 leaf tests each), but gsd-test-summary compares results by title and reported 29+29 false "only in Mac / only in Docker" discrepancies for tests that actually pass everywhere. Move the golden path out of the title and into the assertion message (still shown on failure); titles are now byte-identical across platforms so the cross-platform comparator matches them. No assertion logic or golden values changed. Refs #1169. Co-Authored-By: Claude Opus 4.8 <noreply@anthropic.com> * fix(#1169): derive TDD_MODE via gsd_run --active-cap, not node -e (fix prompt-injection CI gate) The prior fix reverted execute-phase.md:181 from jq to `node -e` to close a Codex HIGH (jq||echo-false silently disabling the MVP+TDD blocking gate on jq-less runtimes) — but the CI prompt-injection scanner BLOCKS new `node -e` in workflow markdown (inline code-exec = injection vector), turning the security gate red. Both forms were wrong: node -e fails the scanner; jq fail-opens a blocking safety gate; `config-get workflow.tdd_mode` is forbidden by the conformance leak gate (tdd_mode is capability-owned). Correct fix (what Codex recommended): a gsd_run-native boolean. Add an `--active-cap <capId>` flag to `loop render-hooks <point>` that resolves hooks the normal way and prints exactly `true`/`false` for whether a capId is active — scanner-safe (canonical launcher, no inline code), node-reliable (no optional jq to fail-open), and leak-free (render-hooks resolution, not config-get). execute-phase.md:181 now `TDD_MODE=$(gsd_run loop render-hooks execute:post --active-cap tdd)`. +5 behavioral tests for the flag. Verified: prompt-injection-scan --diff origin/next → 0 findings; conformance gate 13/13 (execute-phase.md 92934 < 93166); execute-mvp-tdd + tdd-mode + loop-render-hooks 87/0; lint 0/0. Refs #1167, #1169. Co-Authored-By: Claude Opus 4.8 <noreply@anthropic.com> --------- Co-authored-by: Claude Opus 4.8 <noreply@anthropic.com>
464 lines
20 KiB
JavaScript
464 lines
20 KiB
JavaScript
'use strict';
|
|
|
|
/**
|
|
* E2E content tests for execute:post hook resolution and tdd.review-checkpoint gate.
|
|
*
|
|
* Hook point: execute:post
|
|
* Focus:
|
|
* - loop render-hooks execute:post typed envelope (step + gate ordering, both-on / tdd-off / both-off)
|
|
* - check tdd.review-checkpoint via CLI subprocess with real git fixtures:
|
|
* RED+GREEN → block:false,violations:0,Pass
|
|
* no commits → block:true,missing:[RED,GREEN]
|
|
* RED only → block:true,missing:[GREEN]
|
|
* no type:tdd plans → block:false,tddPlans:0
|
|
* violations=1 boundary → block:true with advisory table
|
|
* missing phase arg → exitCode:1
|
|
* - rendered text format: Step 1 code-review before Gate tdd
|
|
*
|
|
* HARD RULES followed:
|
|
* - CONTENT/E2E only: every test drives a real CLI subprocess or real resolver
|
|
* - No readFileSync source-grep (scripts/lint-no-source-grep.cjs would reject it)
|
|
* - Genuine assertions: negative/BVA cases assert the SPECIFIC differing value
|
|
* - Fully isolated: each test has its own createTempProject / createTempGitProject
|
|
* - Git fixtures use real file commits (not --allow-empty) so git log --grep -- . matches
|
|
*/
|
|
|
|
const { describe, test, afterEach } = require('node:test');
|
|
const assert = require('node:assert/strict');
|
|
const fs = require('node:fs');
|
|
const os = require('node:os');
|
|
const path = require('node:path');
|
|
const { execFileSync, spawnSync } = require('node:child_process');
|
|
|
|
const { cleanup } = require('./helpers.cjs');
|
|
|
|
const TOOLS_PATH = path.join(__dirname, '..', 'gsd-core', 'bin', 'gsd-tools.cjs');
|
|
|
|
// ─── Git fixture helper (inlined — do NOT modify helpers.cjs) ──────────────────
|
|
|
|
/**
|
|
* Create a temp dir with a git repo and initial commit containing a .planning/
|
|
* phases directory structure. Commits real files (not --allow-empty) so that
|
|
* git log --grep -- . works correctly (the -- path filter skips empty-tree commits).
|
|
*
|
|
* Returns { tmpDir } — cleanup() in afterEach.
|
|
*/
|
|
function createTddGitFixture({ planFiles = [] } = {}) {
|
|
const tmpDir = fs.mkdtempSync(path.join(os.tmpdir(), 'gsd-tdd-e2e-'));
|
|
|
|
function git(...args) {
|
|
const result = spawnSync('git', args, {
|
|
cwd: tmpDir,
|
|
encoding: 'utf-8',
|
|
env: {
|
|
...process.env,
|
|
GIT_AUTHOR_NAME: 'Test',
|
|
GIT_AUTHOR_EMAIL: 'test@test.com',
|
|
GIT_COMMITTER_NAME: 'Test',
|
|
GIT_COMMITTER_EMAIL: 'test@test.com',
|
|
},
|
|
});
|
|
if (result.status !== 0) {
|
|
throw new Error(`git ${args.join(' ')} failed: ${result.stderr}`);
|
|
}
|
|
return result.stdout.trim();
|
|
}
|
|
|
|
git('init', '--initial-branch=main');
|
|
git('config', 'user.email', 'test@test.com');
|
|
git('config', 'user.name', 'Test');
|
|
|
|
// Create planning directory
|
|
const planningDir = path.join(tmpDir, '.planning');
|
|
const phasesDir = path.join(planningDir, 'phases');
|
|
fs.mkdirSync(planningDir, { recursive: true });
|
|
|
|
// Write plan files
|
|
for (const { dir, filename, content } of planFiles) {
|
|
const phaseDir = path.join(phasesDir, dir);
|
|
fs.mkdirSync(phaseDir, { recursive: true });
|
|
fs.writeFileSync(path.join(phaseDir, filename), content, 'utf8');
|
|
}
|
|
|
|
// Write a config.json with git tracking so initial commit has a real file
|
|
fs.writeFileSync(path.join(planningDir, 'config.json'), '{}', 'utf8');
|
|
|
|
git('add', '.');
|
|
git('commit', '-m', 'init: project scaffold');
|
|
|
|
return { tmpDir, git };
|
|
}
|
|
|
|
/**
|
|
* Build a type:tdd PLAN.md frontmatter block content.
|
|
*/
|
|
function tddPlan(phaseNum, planId) {
|
|
return `---\ntype: tdd\nphase: ${phaseNum}\nslug: ${planId}\n---\n# Task: ${planId}\n`;
|
|
}
|
|
|
|
/**
|
|
* Build a type:execute PLAN.md (non-TDD) content.
|
|
*/
|
|
function executePlan(phaseNum, planId) {
|
|
return `---\ntype: execute\nphase: ${phaseNum}\nslug: ${planId}\n---\n# Task: ${planId}\n`;
|
|
}
|
|
|
|
/**
|
|
* Commit a real file in the git fixture with the given commit message.
|
|
* Needed because git log --grep with -- path filter only matches commits
|
|
* that changed at least one tracked file.
|
|
*/
|
|
function commitFile(git, tmpDir, filename, commitMessage) {
|
|
const filepath = path.join(tmpDir, filename);
|
|
// Append timestamp to make each file unique
|
|
fs.writeFileSync(filepath, `${commitMessage}\n${Date.now()}\n`, 'utf8');
|
|
git('add', filepath);
|
|
git('commit', '-m', commitMessage);
|
|
}
|
|
|
|
// ─── Helpers for subprocess invocation ────────────────────────────────────────
|
|
|
|
const TEST_ENV_BASE = {
|
|
GSD_SESSION_KEY: '',
|
|
CODEX_THREAD_ID: '',
|
|
CLAUDE_SESSION_ID: '',
|
|
CLAUDE_CODE_SSE_PORT: '',
|
|
OPENCODE_SESSION_ID: '',
|
|
GEMINI_SESSION_ID: '',
|
|
CURSOR_SESSION_ID: '',
|
|
WINDSURF_SESSION_ID: '',
|
|
TERM_SESSION_ID: '',
|
|
WT_SESSION: '',
|
|
TMUX_PANE: '',
|
|
ZELLIJ_SESSION_NAME: '',
|
|
TTY: '',
|
|
SSH_TTY: '',
|
|
};
|
|
|
|
function runTools(args, cwd) {
|
|
const argv = Array.isArray(args)
|
|
? args
|
|
: (args.match(/(?:[^\s"']+|"[^"]*"|'[^']*')+/g) || [])
|
|
.map((t) => t.replace(/"([^"]*)"/g, '$1').replace(/'([^']*)'/g, '$1'));
|
|
|
|
try {
|
|
const stdout = execFileSync(process.execPath, [TOOLS_PATH, ...argv], {
|
|
cwd,
|
|
encoding: 'utf-8',
|
|
env: { ...process.env, ...TEST_ENV_BASE },
|
|
timeout: 60000,
|
|
});
|
|
return { success: true, output: stdout.trim(), exitCode: 0, error: '' };
|
|
} catch (err) {
|
|
return {
|
|
success: false,
|
|
output: err.stdout?.toString().trim() || '',
|
|
error: err.stderr?.toString().trim() || err.message,
|
|
exitCode: err.status ?? 1,
|
|
};
|
|
}
|
|
}
|
|
|
|
// ─── Tests ─────────────────────────────────────────────────────────────────────
|
|
|
|
describe('execute:post render-hooks — typed envelope resolution', () => {
|
|
let tmpDir;
|
|
|
|
afterEach(() => { if (tmpDir) { cleanup(tmpDir); tmpDir = null; } });
|
|
|
|
test('[happy] tdd_mode=true and code_review=true: both hooks in typed shape with step before gate', () => {
|
|
tmpDir = fs.mkdtempSync(path.join(os.tmpdir(), 'gsd-ep-'));
|
|
fs.mkdirSync(path.join(tmpDir, '.planning'), { recursive: true });
|
|
fs.writeFileSync(
|
|
path.join(tmpDir, '.planning', 'config.json'),
|
|
JSON.stringify({ workflow: { code_review: true, tdd_mode: true } }),
|
|
'utf8'
|
|
);
|
|
|
|
const result = runTools('loop render-hooks execute:post --raw', tmpDir);
|
|
assert.ok(result.success, `render-hooks should succeed. stderr: ${result.error}`);
|
|
|
|
const envelope = JSON.parse(result.output);
|
|
assert.strictEqual(envelope.point, 'execute:post', 'point field must be execute:post');
|
|
assert.ok(Array.isArray(envelope.activeHooks), 'activeHooks must be an array');
|
|
assert.strictEqual(envelope.activeHooks.length, 2, 'both hooks (step + gate) must be active');
|
|
|
|
// Step 1: code-review step (must come before gate)
|
|
const step = envelope.activeHooks[0];
|
|
assert.strictEqual(step.kind, 'step', 'first hook must be a step');
|
|
assert.strictEqual(step.capId, 'code-review', 'step capId must be code-review');
|
|
assert.deepStrictEqual(step.ref, { skill: 'code-review' }, 'step ref must point to code-review skill');
|
|
assert.ok(Array.isArray(step.produces), 'produces must be array');
|
|
assert.ok(step.produces.includes('REVIEW.md'), 'step must produce REVIEW.md');
|
|
assert.strictEqual(step.onError, 'skip', 'code-review step onError must be skip');
|
|
|
|
// Gate: tdd advisory gate (must come after step)
|
|
const gate = envelope.activeHooks[1];
|
|
assert.strictEqual(gate.kind, 'gate', 'second hook must be a gate');
|
|
assert.strictEqual(gate.capId, 'tdd', 'gate capId must be tdd');
|
|
assert.deepStrictEqual(gate.check, { query: 'tdd.review-checkpoint' }, 'gate check query must match');
|
|
assert.strictEqual(gate.blocking, false, 'tdd gate must be advisory (blocking=false)');
|
|
});
|
|
|
|
test('[negative] code_review=false and tdd_mode=false: empty activeHooks with no-hooks rendered text', () => {
|
|
tmpDir = fs.mkdtempSync(path.join(os.tmpdir(), 'gsd-ep-'));
|
|
fs.mkdirSync(path.join(tmpDir, '.planning'), { recursive: true });
|
|
fs.writeFileSync(
|
|
path.join(tmpDir, '.planning', 'config.json'),
|
|
JSON.stringify({ workflow: { code_review: false, tdd_mode: false } }),
|
|
'utf8'
|
|
);
|
|
|
|
const result = runTools('loop render-hooks execute:post --raw', tmpDir);
|
|
assert.ok(result.success, `render-hooks should succeed even with both disabled. stderr: ${result.error}`);
|
|
|
|
const envelope = JSON.parse(result.output);
|
|
assert.strictEqual(envelope.point, 'execute:post');
|
|
// SPECIFIC assertion: 0 hooks, not 1 or 2
|
|
assert.strictEqual(envelope.activeHooks.length, 0, 'both disabled: must return ZERO active hooks, not any');
|
|
assert.ok(
|
|
envelope.rendered.includes('_No active hooks at execute:post._'),
|
|
`rendered must contain placeholder text, got: ${envelope.rendered}`
|
|
);
|
|
});
|
|
|
|
test('[negative] tdd_mode=false excludes tdd gate but code-review step active by schema default', () => {
|
|
tmpDir = fs.mkdtempSync(path.join(os.tmpdir(), 'gsd-ep-'));
|
|
fs.mkdirSync(path.join(tmpDir, '.planning'), { recursive: true });
|
|
fs.writeFileSync(
|
|
path.join(tmpDir, '.planning', 'config.json'),
|
|
JSON.stringify({ workflow: { tdd_mode: false } }),
|
|
'utf8'
|
|
);
|
|
|
|
const result = runTools('loop render-hooks execute:post --raw', tmpDir);
|
|
assert.ok(result.success, `render-hooks should succeed. stderr: ${result.error}`);
|
|
|
|
const envelope = JSON.parse(result.output);
|
|
// SPECIFIC assertion: exactly 1 hook (code-review only), not 0 or 2
|
|
assert.strictEqual(envelope.activeHooks.length, 1, 'tdd_mode=false: exactly 1 hook (step only), not 2');
|
|
assert.strictEqual(envelope.activeHooks[0].capId, 'code-review', 'sole hook must be code-review step');
|
|
assert.strictEqual(envelope.activeHooks[0].kind, 'step', 'sole hook must be kind=step');
|
|
// Confirm tdd gate is absent
|
|
const tddHook = envelope.activeHooks.find((h) => h.capId === 'tdd');
|
|
assert.strictEqual(tddHook, undefined, 'no tdd gate hook must be present when tdd_mode=false');
|
|
});
|
|
|
|
test('[happy] rendered text format: Step 1 code-review before Gate tdd in correct markdown', () => {
|
|
tmpDir = fs.mkdtempSync(path.join(os.tmpdir(), 'gsd-ep-'));
|
|
fs.mkdirSync(path.join(tmpDir, '.planning'), { recursive: true });
|
|
fs.writeFileSync(
|
|
path.join(tmpDir, '.planning', 'config.json'),
|
|
JSON.stringify({ workflow: { code_review: true, tdd_mode: true } }),
|
|
'utf8'
|
|
);
|
|
|
|
const result = runTools('loop render-hooks execute:post --raw', tmpDir);
|
|
assert.ok(result.success, `render-hooks should succeed. stderr: ${result.error}`);
|
|
|
|
const envelope = JSON.parse(result.output);
|
|
const rendered = envelope.rendered;
|
|
|
|
// Step 1 code-review heading must appear first
|
|
assert.ok(
|
|
rendered.includes('### Step 1: skill:code-review (code-review)'),
|
|
`rendered must start with Step 1 heading. got: ${rendered.slice(0, 200)}`
|
|
);
|
|
// produces and consumes in step section
|
|
assert.ok(rendered.includes('produces: REVIEW.md'), 'rendered must include produces: REVIEW.md');
|
|
assert.ok(rendered.includes('consumes: SUMMARY.md'), 'rendered must include consumes: SUMMARY.md');
|
|
// when key for code-review step
|
|
assert.ok(rendered.includes('when: `workflow.code_review`'), 'rendered must include when for code-review');
|
|
// Gate tdd appears AFTER the step
|
|
assert.ok(
|
|
rendered.includes('**Gate** (tdd): check={"query":"tdd.review-checkpoint"}, blocking=false, onError=skip'),
|
|
`rendered must include Gate tdd section. got: ${rendered}`
|
|
);
|
|
// Step 1 must come before the gate
|
|
const step1Idx = rendered.indexOf('### Step 1');
|
|
const gateIdx = rendered.indexOf('**Gate** (tdd)');
|
|
assert.ok(step1Idx < gateIdx, 'Step 1 code-review must appear before Gate tdd in rendered text');
|
|
});
|
|
});
|
|
|
|
// ─── check tdd.review-checkpoint via CLI — git fixture tests ───────────────────
|
|
|
|
describe('check tdd.review-checkpoint — CLI subprocess E2E with git fixtures', () => {
|
|
test('[happy] RED+GREEN commits present: block:false, violations:0, status Pass', () => {
|
|
const { tmpDir, git } = createTddGitFixture({
|
|
planFiles: [
|
|
{ dir: '01-phase1', filename: '01-01-PLAN.md', content: tddPlan(1, '01-01') },
|
|
],
|
|
});
|
|
|
|
try {
|
|
// RED: failing test commit (must touch a real file for git log --grep -- . to work)
|
|
commitFile(git, tmpDir, 'test-login.js', 'test(01-01): failing test for login');
|
|
// GREEN: implementation commit
|
|
commitFile(git, tmpDir, 'login.js', 'feat(01-01): implement login');
|
|
|
|
const result = runTools('check tdd.review-checkpoint 1 --raw', tmpDir);
|
|
assert.ok(result.success, `check should succeed with exit 0. stderr: ${result.error}`);
|
|
|
|
const out = JSON.parse(result.output);
|
|
assert.strictEqual(out.block, false, 'RED+GREEN present: block must be false, not true');
|
|
assert.strictEqual(out.passed, true, 'passed must be true');
|
|
assert.strictEqual(out.tddPlans, 1, 'must find 1 tdd plan');
|
|
assert.strictEqual(out.violations, 0, 'violations must be 0 when both commits present');
|
|
assert.ok(Array.isArray(out.rows), 'rows must be array');
|
|
assert.strictEqual(out.rows.length, 1, 'must have 1 row');
|
|
assert.strictEqual(out.rows[0].planId, '01-01', 'planId must be 01-01');
|
|
assert.strictEqual(out.rows[0].red, true, 'red must be true');
|
|
assert.strictEqual(out.rows[0].green, true, 'green must be true');
|
|
assert.strictEqual(out.rows[0].status, 'Pass', 'status must be Pass');
|
|
assert.strictEqual(out.rows[0].missing.length, 0, 'missing array must be empty');
|
|
} finally {
|
|
cleanup(tmpDir);
|
|
}
|
|
});
|
|
|
|
test('[negative] type:tdd plan with no commits: block:true, violations:1, missing includes RED and GREEN', () => {
|
|
const { tmpDir } = createTddGitFixture({
|
|
planFiles: [
|
|
{ dir: '01-phase1', filename: '01-01-PLAN.md', content: tddPlan(1, '01-01') },
|
|
],
|
|
});
|
|
|
|
try {
|
|
// No additional commits — only the init commit exists
|
|
|
|
const result = runTools('check tdd.review-checkpoint 1 --raw', tmpDir);
|
|
assert.ok(result.success, `check should exit 0 (advisory gate). stderr: ${result.error}`);
|
|
|
|
const out = JSON.parse(result.output);
|
|
// SPECIFIC assertion: block must be TRUE (distinguishes from the passing case)
|
|
assert.strictEqual(out.block, true, 'no commits: block must be TRUE, not false');
|
|
assert.strictEqual(out.tddPlans, 1, 'must find 1 tdd plan');
|
|
assert.strictEqual(out.violations, 1, 'violations must be 1');
|
|
assert.strictEqual(out.rows[0].red, false, 'red must be false without test() commit');
|
|
assert.strictEqual(out.rows[0].green, false, 'green must be false without feat() commit');
|
|
assert.strictEqual(out.rows[0].status, 'FAIL', 'status must be FAIL');
|
|
// missing must include both RED and GREEN
|
|
assert.ok(out.rows[0].missing.includes('RED'), 'missing must include RED');
|
|
assert.ok(out.rows[0].missing.includes('GREEN'), 'missing must include GREEN');
|
|
} finally {
|
|
cleanup(tmpDir);
|
|
}
|
|
});
|
|
|
|
test('[negative] RED present but GREEN missing: block:true, violations:1, missing deepEqual [GREEN]', () => {
|
|
const { tmpDir, git } = createTddGitFixture({
|
|
planFiles: [
|
|
{ dir: '01-phase1', filename: '01-01-PLAN.md', content: tddPlan(1, '01-01') },
|
|
],
|
|
});
|
|
|
|
try {
|
|
// Only RED commit — no feat() commit
|
|
commitFile(git, tmpDir, 'test-auth.js', 'test(01-01): failing auth test');
|
|
|
|
const result = runTools('check tdd.review-checkpoint 1 --raw', tmpDir);
|
|
assert.ok(result.success, `check should exit 0. stderr: ${result.error}`);
|
|
|
|
const out = JSON.parse(result.output);
|
|
// SPECIFIC assertion: block true, violations 1
|
|
assert.strictEqual(out.block, true, 'RED only: block must be true');
|
|
assert.strictEqual(out.violations, 1, 'violations must be exactly 1');
|
|
assert.strictEqual(out.rows[0].red, true, 'red must be true (commit present)');
|
|
assert.strictEqual(out.rows[0].green, false, 'green must be false (no feat commit)');
|
|
assert.strictEqual(out.rows[0].status, 'FAIL', 'status must be FAIL');
|
|
// missing must be exactly ['GREEN'] — not ['RED', 'GREEN']
|
|
assert.deepStrictEqual(out.rows[0].missing, ['GREEN'], 'missing must deepEqual [GREEN] when RED present');
|
|
} finally {
|
|
cleanup(tmpDir);
|
|
}
|
|
});
|
|
|
|
test('[empty-resolution] no type:tdd plans (type:execute only): block:false, tddPlans:0, empty rows', () => {
|
|
const { tmpDir } = createTddGitFixture({
|
|
planFiles: [
|
|
{ dir: '01-phase1', filename: '01-01-PLAN.md', content: executePlan(1, '01-01') },
|
|
],
|
|
});
|
|
|
|
try {
|
|
const result = runTools('check tdd.review-checkpoint 1 --raw', tmpDir);
|
|
assert.ok(result.success, `check should succeed with exit 0. stderr: ${result.error}`);
|
|
|
|
const out = JSON.parse(result.output);
|
|
// SPECIFIC assertion: block false AND tddPlans 0 (distinguishes from a plan that passes)
|
|
assert.strictEqual(out.block, false, 'no tdd plans: block must be false');
|
|
assert.strictEqual(out.tddPlans, 0, 'tddPlans must be 0 when no type:tdd files');
|
|
assert.strictEqual(out.violations, 0, 'violations must be 0');
|
|
assert.strictEqual(out.rows.length, 0, 'rows must be empty array');
|
|
assert.strictEqual(out.table, '', 'table must be empty string when no tdd plans');
|
|
} finally {
|
|
cleanup(tmpDir);
|
|
}
|
|
});
|
|
|
|
test('[bva] violations=1 boundary: exactly 1 violation sets block:true and advisory message present', () => {
|
|
// Two tdd plans: 01-01 passes (both commits), 01-02 fails (no commits)
|
|
// violations = 1 exactly — boundary test (violations > 0 → block:true)
|
|
const { tmpDir, git } = createTddGitFixture({
|
|
planFiles: [
|
|
{ dir: '01-phase1', filename: '01-01-PLAN.md', content: tddPlan(1, '01-01') },
|
|
{ dir: '01-phase1', filename: '01-02-PLAN.md', content: tddPlan(1, '01-02') },
|
|
],
|
|
});
|
|
|
|
try {
|
|
// 01-01: both RED and GREEN commits (passes)
|
|
commitFile(git, tmpDir, 'test1.js', 'test(01-01): failing test');
|
|
commitFile(git, tmpDir, 'impl1.js', 'feat(01-01): implementation');
|
|
// 01-02: no commits (fails)
|
|
|
|
const result = runTools('check tdd.review-checkpoint 1 --raw', tmpDir);
|
|
assert.ok(result.success, `check should exit 0. stderr: ${result.error}`);
|
|
|
|
const out = JSON.parse(result.output);
|
|
// SPECIFIC: block must be TRUE for violations=1 (not false as it would be for violations=0)
|
|
assert.strictEqual(out.block, true, 'violations=1 boundary: block must be true');
|
|
assert.strictEqual(out.violations, 1, 'violations must be exactly 1 (not 0, not 2)');
|
|
assert.strictEqual(out.tddPlans, 2, 'tddPlans must be 2');
|
|
assert.strictEqual(out.passed, true, 'advisory gate: passed stays true');
|
|
|
|
// Check both rows
|
|
const passRow = out.rows.find((r) => r.planId === '01-01');
|
|
const failRow = out.rows.find((r) => r.planId === '01-02');
|
|
assert.ok(passRow, '01-01 row must exist');
|
|
assert.ok(failRow, '01-02 row must exist');
|
|
assert.strictEqual(passRow.status, 'Pass', '01-01 must Pass');
|
|
assert.strictEqual(failRow.status, 'FAIL', '01-02 must FAIL');
|
|
|
|
// Advisory table must mention the warning text
|
|
assert.ok(
|
|
out.table.includes('⚠ Gate violations are advisory'),
|
|
'table must include advisory warning when violations > 0'
|
|
);
|
|
} finally {
|
|
cleanup(tmpDir);
|
|
}
|
|
});
|
|
|
|
test('[negative] missing phase argument: exitCode 1 and error contains required message', () => {
|
|
const tmpDir = fs.mkdtempSync(path.join(os.tmpdir(), 'gsd-tdd-noarg-'));
|
|
fs.mkdirSync(path.join(tmpDir, '.planning'), { recursive: true });
|
|
|
|
try {
|
|
const result = runTools('check tdd.review-checkpoint --raw', tmpDir);
|
|
// SPECIFIC: exitCode must be 1 (non-zero), not 0
|
|
assert.strictEqual(result.success, false, 'missing phase arg must cause failure (success=false)');
|
|
assert.strictEqual(result.exitCode, 1, 'exitCode must be 1, not 0');
|
|
// Error message must identify the command and what's missing
|
|
const errText = result.error + result.output;
|
|
assert.ok(
|
|
errText.includes('tdd.review-checkpoint') || errText.includes('phase argument'),
|
|
`error must reference tdd.review-checkpoint or phase argument. got: ${errText}`
|
|
);
|
|
} finally {
|
|
cleanup(tmpDir);
|
|
}
|
|
});
|
|
});
|