Files
msd-core/tests/loop-hooks-verify-post-e2e.test.cjs
Tom Boucher b10e56818b feat(#1169): complete ADR-857 phase 6 — migrate features to Capabilities, revive dead gates, harden conformance gate (#1183)
* test(#1168): make phase-6 gate un-gameable — reject empty stubs + require loop shrink

The migration assertion previously checked only role==feature, so a registration-only stub (empty hooks, logic left inline) would turn the gate green while phase 6 stayed incomplete — the exact false-completion pattern this gate exists to prevent. Strengthen it: each ADR-named feature must OWN its behavior (>=1 hook, or a command family); and plan-phase.md/execute-phase.md must shrink strictly below their frozen pre-phase-6 sizes (94519/93166 LF bytes), which also defeats double-run gaming (declare a hook but keep the inline block -> file does not shrink -> red).

Gate now 5 pass / 4 fail (orphaned execute:wave:post, empty/unregistered features, config-key leaks, no shrink). Green is now reachable only by REAL migration. Refs #1168, #1169.

Co-Authored-By: Claude Opus 4.8 <noreply@anthropic.com>

* feat(#1169): migrate gap-analysis to a Capability (plan:post gate)

First real ADR-857 phase-6 migration (pattern-defining tracer). gap-analysis moves from an inline post_planning_gaps branch in plan-phase.md to a real plan:post gate Capability:

- capabilities/gap-analysis/capability.json: role:feature, plan:post gate (when=workflow.post_planning_gaps, blocking:false advisory), OWNS workflow.post_planning_gaps (federated out of central schema). - plan-phase.md: inline config-get + gsd_run gap-analysis block replaced with a plan:post render-hooks call site dispatching the gate; file shrinks 94519->93279. - src/check-command-router.cts: cmdGapAnalysisPlanPost runs the real gap analysis via gap-checker. - post_planning_gaps removed from central manifest; resolves via federated config (default true preserved). - tests/post-planning-gaps-2493: re-pointed to assert capability ownership.

Verified: gate 5 pass / 4 fail (gap-analysis cleared from migration, plan:post-orphan, config-leak, and plan-phase shrink checks); loadConfig still returns post_planning_gaps=true; check command runs real analysis; 392/392 in the config/registry/federation/router net. Refs #1169.

Co-Authored-By: Claude Opus 4.8 <noreply@anthropic.com>

* feat(#1169): migrate profile-pipeline to a command-family Capability

ADR-857 Decision 7: profile-pipeline becomes a command-family Capability (like audit/intel/graphify). capabilities/profile-pipeline/capability.json declares an 8-command family (scan-sessions, extract-messages, profile-sample, write-profile, profile-questionnaire, generate-dev-preferences, generate-claude-profile, generate-claude-md) backed by a new gsd-core/bin/lib/profile-pipeline-command-router.cjs; the inline case arms are removed from gsd-tools.cjs. Owns profile-pipeline.enabled (federated).

Verified: registry shows role:feature with commands.length=8; scan-sessions/profile-sample run live via the family; gate cleared profile-pipeline from the empty-stub failure (only tdd/schema-gate/drift remain); 296/296 registry+inventory+gsd-tools tests; lint 0 errors. Refs #1169.

Co-Authored-By: Claude Opus 4.8 <noreply@anthropic.com>

* fix(#1167): wire execute:wave:post + implement ui.safety-gate check

Revives the second dead gate from #1167: ui.gates@execute:wave:post was declared but never dispatched AND its check.query (ui.safety-gate) was unimplemented. Adds the per-wave execute:wave:post render-hooks call site in execute-phase.md (fires after each wave's merge/cleanup, before the next forks) and implements cmdUiSafetyGate (frontend + UI-SPEC aware, mirrors cmdUiPlanGate) in check-command-router. +17 regression tests.

Verified: phase-6 orphaned-points conformance test now PASSES (gate 6 pass / 3 fail); ui-safety-gate routable in dot+hyphen forms; check-ui-safety-gate 17/17, check-ui-plan-gate 18/18; lint 0 errors. Refs #1167, #1168.

Co-Authored-By: Claude Opus 4.8 <noreply@anthropic.com>

* feat(#1169): migrate drift (schema + codebase) to execute:wave:post gates

Removes the inline schema_drift_gate + codebase_drift_gate steps (77 lines) from execute-phase.md; drift becomes a Capability with two execute:wave:post gates (verify.schema-drift blocking, verify.codebase-drift advisory) dispatched via the per-wave render-hooks call site. check-command-router routes verify.schema-drift / verify.codebase-drift to the real detectors. Federates workflow.drift_threshold / drift_action / schema_drift_gate out of central.

Also fixes the execute:wave:post dispatch prose to run NON-blocking (advisory) gates too — the prior version only ran blocking gates, which would have silently dropped the codebase-drift advisory after its inline step was removed. Behavior preserved.

Verified: gate 7 pass / 2 fail (drift cleared from stub + config-leak; execute-phase.md 92297 < 93166 frozen -> shrink passes); both drift checks run real detection; loadConfig defaults preserved (threshold=3, action=warn, gate=true); drift-detection 56/56 + schema-drift 34/34; lint 0 errors. Refs #1169.

Co-Authored-By: Claude Opus 4.8 <noreply@anthropic.com>

* feat(#1169): migrate tdd to a Capability (plan:pre contribution + execute:post gate)

tdd becomes a real Capability: a plan:pre contribution injects the <tdd_mode_active> planner guidance (rendered from PLAN_PRE_HOOKS_JSON like security's contribution), and an execute:post gate (tdd.review-checkpoint, advisory) runs the real end-of-phase RED/GREEN review via a new check-command handler. Inline tdd_mode reads + the inline planner block + the tdd_review_checkpoint step are removed; workflow.tdd_mode is federated out of central. The MVP+TDD per-task RED-commit gate is preserved — TDD_MODE is now derived from the execute:post hooks (capId==tdd active), not an inline config-get.

BEHAVIOR CHANGE (documented, not silent): the --tdd CLI flag now persists workflow.tdd_mode=true via config-set instead of being per-invocation. Rationale: tdd is now a config-toggled Capability, and env vars do not persist across the workflow's separate bash blocks (config does), so an ephemeral override isn't cleanly achievable; --tdd therefore enables the tdd capability, consistent with how all capabilities are toggled.

Verified: gate 7 pass / 2 fail (tdd cleared from stub + config-leak; plan-phase + execute-phase both < frozen sizes); contribution injection + execute:post gate dispatch wired; MVP+TDD gate preserved; tdd.review-checkpoint runs real review; full unit suite 556/0; lint 0 errors. Refs #1169.

Co-Authored-By: Claude Opus 4.8 <noreply@anthropic.com>

* feat(#1169): migrate schema-gate to a plan:pre contribution Capability

The plan-time schema-push detection (former plan-phase.md §5.7) becomes a schema-gate Capability: a plan:pre contribution (into:planner, when:workflow.schema_push_detection) whose fragment carries the full ORM-detection + [BLOCKING] schema-push-task injection logic, rendered into the planner via the existing plan:pre render-hooks dispatch. The inline §5.7 block is removed (plan-phase.md 94519->90445). workflow.schema_push_detection is a new capability-owned (federated) key, default true. (The execute-side schema-drift gate was migrated separately into the drift capability.)

Verified: registry inlines the fragment (len 2704) so it is actually delivered at plan:pre; gate 8 pass / 1 fail — all 5 ADR-named features now real Capabilities, only the config-leak test remains (intel/security, next unit). Refs #1169.

Co-Authored-By: Claude Opus 4.8 <noreply@anthropic.com>

* feat(#1169): close the 3 capability config-key leaks — phase-6 gate now GREEN

Removes the last inline config-get reads of capability-owned keys from plan-phase.md. security_asvs_level/security_block_on now flow through the security plan:pre contribution via a new loop-resolver configValues mechanism (resolves declared config keys with the same 4-level precedence as activation and attaches them to the rendered hook); the §5.55 banner reads them from PLAN_PRE_HOOKS_JSON. intel.enabled becomes a real intel plan:pre step (ref.command: intel api-surface) dispatched via render-hooks; the inline intel branch is gone. gen-capability-registry now validates ref.command as a third dispatch shape.

Verified: phase-6 capstone conformance gate is FULLY GREEN (9/0); 3 leaks gone (grep=0); security configValues resolve to {2,medium}/default {1,high}; intel step present only when enabled; loop-render-hooks 62/0, capability-registry 287/0, capability-state/federated-config 113/0; lint 0 errors. Closes the migration half of #1169. Refs #1139, #1167, #1168.

Co-Authored-By: Claude Opus 4.8 <noreply@anthropic.com>

* fix(#1169): address adversarial review — restore schema-drift block, generic planner injection, uniform gate contract

Adversarial review caught 2 real regressions the green gate missed: (1) schema-drift no longer blocked — the execute:wave:post dispatch read GATE_RESULT.block but verify.schema-drift emitted drift_detected/blocking, and onError:skip wrongly bypassed positive blocks; (2) only tdd's plan:pre contribution was injected into the planner, dropping schema-gate's schema-push detection and security's threat-model guidance.

Fixes: (A) every gate check returns a uniform boolean 'block' under --raw (the dispatch form), with advisory gates (tdd/gap) carrying their report in 'message'; (B) gate-dispatch contract corrected at all sites — onError governs command errors only, a blocking gate's positive block always halts; (C) generic planner injection of all plan:pre contributions where into=='planner' (tdd + schema-gate + security incl configValues); (D) two new conformance assertions: planner contributions injected generically + every gate check.query returns boolean block under --raw.

Verified: gate 11/11; all 6 gate checks return boolean block under --raw; full suite 595/0; lint 0 errors. Refs #1167, #1168, #1169.

Co-Authored-By: Claude Opus 4.8 <noreply@anthropic.com>

* fix(#1169): restore MVP+TDD end-of-phase blocking escalation (2nd adversarial pass)

The migrated tdd execute:post gate is statically blocking:false, but the contract (references/execute-mvp-tdd.md + CONTEXT.md) requires the end-of-phase TDD review to ESCALATE from advisory to blocking when MVP_MODE && TDD_MODE && a TDD plan misses a RED/GREEN commit. The migration prose had downgraded this to a 'strong advisory recommendation' — silent loss of the blocking escalation. Restore it: the tdd-gate dispatch now refuses to mark the phase complete (Phase blocked message) under MVP+TDD when GATE_RESULT.block is true; advisory otherwise.

Also strengthen tests/execute-mvp-tdd-gate.test.cjs: hasBlockingEscalation previously matched any line with 'blocking'+'mvp+tdd' (so 'advisory (blocking: false) ... under MVP+TDD' was a false green); now it requires the real refusal semantics ('refuse to mark the phase complete' / 'phase blocked'). Caught by 2nd adversarial review pass.

Verified: execute-phase.md 92702 < 93166 frozen; mvp-tdd-gate + phase-6 gate 19/0; full suite green; lint 0 errors. Refs #1169.

Co-Authored-By: Claude Opus 4.8 <noreply@anthropic.com>

* fix(#1169): restore MVP+TDD proceed-block, codebase auto-remap, schema skip-flag (3rd adversarial pass)

3rd adversarial pass found 4 more silent regressions: (1) the tdd MVP+TDD 'refuse to mark complete' was nullified by a downstream 'ALWAYS proceed regardless of gate results' line — proceed is now conditional (stops on an active MVP+TDD block); (2) the test now asserts the proceed is NOT an unconditional override; (3) codebase-drift auto-remap (spawn gsd-codebase-mapper when drift_action=auto-remap) was dropped — the execute:wave:post advisory dispatch now consumes spawn_mapper/directive; (4) GSD_SKIP_SCHEMA_CHECK bypass was lost from the gate path — cmdVerifySchemaDrift now honors the env var (block:false when set).

Verified: no unconditional proceed; GSD_SKIP_SCHEMA_CHECK=true -> block:false; gate 11/11 + mvp-tdd 9/9; full suite 569/0; lint 0; execute-phase.md 93109 < 93166. Refs #1169.

Co-Authored-By: Claude Opus 4.8 <noreply@anthropic.com>

* fix(#1169): init.cts reads federated config keys from nested path (4th adversarial pass)

Config federation moved tdd_mode/research/nyquist_validation from flat config.<key> to nested config.workflow.<key>, but src/init.cts still read them flat — so init.plan-phase/init.execute-phase emitted tdd_mode:false / research_enabled:undefined / nyquist:undefined regardless of config (a public command-contract regression; the migrated loops use render-hooks so enforcement was unaffected). Read via config.workflow (type-safe Record cast). Now init reflects the same resolved values + federated defaults (research/nyquist default true) as the render-hooks path.

Verified: build clean; init.plan-phase emits tdd_mode:true/research:false/nyquist:false for set config, defaults true for empty; full suite 591/0; lint 0. Refs #1169.

Co-Authored-By: Claude Opus 4.8 <noreply@anthropic.com>

* docs(#1169): add changeset for ADR-857 phase-6 completion (PR #1183)

Co-Authored-By: Claude Opus 4.8 <noreply@anthropic.com>

* fix(#1169): complete phase-6 migration fallout — restore TEXT_MODE, fix registry .claude leak, re-point stale workflow-contract tests

The capability migration left real regressions and stale consumer tests that
the per-module unit suite missed but the full cross-platform suite caught (27
failing tests):

Real source regressions (fixed):
- execute-phase.md lost its AskUserQuestion TEXT_MODE plain-text fallback when
  the inline schema_drift_gate step was removed — non-Claude runtimes would
  stall. Restored, and the execute:post gate-dispatch prose de-duplicated to
  cite the execute:wave:post contract (loop body shrinks below the frozen
  pre-phase-6 ceiling while keeping every onError/blocking nuance).
- capabilities/tdd inline fragment hardcoded `@~/.claude/gsd-core/references/tdd.md`,
  baked verbatim into the committed capability-registry.cjs and leaked the
  install path on 11 non-Claude runtimes (registry .cjs is copied, not
  path-converted). Made the fragment path-free; regenerated the registry. The
  phase-6 conformance gate now guards this (no ~/.claude install path in any
  capability source or the generated registry).
- plan-phase.md: removed a §5.7 stub re-added in error and routed Branch 2 to
  step 6 (schema-gate is a plan:pre capability, §5.7 is gone).

Stale workflow-contract tests re-pointed to the capability dispatch they now
must assert (behavior verified preserved in source first, assertions kept
equal-or-stronger): bug-621 + bug-2851 (gap-analysis via gsd_run render-hooks
plan:post + registry binding), feat-2527 (tdd_mode federated out of central),
phase6-planning + plan-phase-ui-redirect (§5.6 bounded by ## 6.),
plan-phase-drift-guard (intel when:intel.enabled skip branch).

profile-pipeline-command-router.cjs un-ignored from eslint (hand-written, no
TS source) + stale disable comments removed. Size baseline regenerated.

Verified: full suite 15140 tests / 0 fail; lint 0 errors; conformance gate green
legitimately. Refs #1139, #1167, #1168, #1169.

Co-Authored-By: Claude Opus 4.8 <noreply@anthropic.com>

* test(#1169): add ADR-857 E2E content-test coverage for the 12 loop points + capability deliverables

Grounds the capability engine in behavioral E2E tests (drive the real
render-hooks/check CLI + the real registry, assert typed result content — no
source-grep), structured around what ADR-857 says to deliver. 207 tests; each
genuineness-checked (flip the expectation, confirm it fails).

Per-loop-point dispatch (7 files): empty-point negative-space across the 6
no-hook points; verify:post 3-step resolution+ordering+onError; plan:pre
contribution/configValues + ui.plan-gate + intel; plan:post gap-analysis;
execute:wave:post drift+ui gates via the check route (schema-drift block/skip,
codebase-drift threshold BVA, auto-remap); execute:post tdd.review-checkpoint
RED/GREEN; ship:pre security gate resolution + frontmatter-get predicate pieces.

ADR-deliverable coverage (4 files): predicate boundary held (edge/prohibition
probes stay core, not off-by-default Feature Capabilities — phase-6 exception);
core loop runs with zero capabilities (all 12 points empty, init bundles
resolve); contribution merge (multiple ordered <contribution from=> blocks);
federated-config key removal on uninstall.

federated-config allowlisted for its 3-file split (unit + integration +
lifecycle). Refs #1139, #1167, #1168, #1169.

Co-Authored-By: Claude Opus 4.8 <noreply@anthropic.com>

* fix(#1169): remove dead drifted converter dups + address adversarial review

Lint cleanup (root-caused, not waved off): src/runtime-artifact-conversion.cts
carried 11 agent-converter functions (+5 orphaned consts/helpers) that were
never exported, never called, and had silently DRIFTED from the live
hand-authored copies in bin/install.js (one even referenced an undefined
`claudeToCopilotTools`). Deleted the dead duplicates; install.js's live copies
are untouched (it never imported these). Lint now 0 errors / 0 warnings.

Adversarial-review (Codex) findings fixed:
- HIGH: execute-phase.md TDD_MODE used `jq ... || echo false`, silently
  disabling the MVP+TDD blocking gate on jq-less runtimes. Reverted to the
  `node -e` form (node is guaranteed; matches the file's other node-e usages) so
  a missing optional tool can no longer fail-open a blocking safety path.
- MEDIUM: federated-config-key-removal orphan-key test was vacuous (it skipped
  the orphan assertion). Now asserts the removed capability's key is genuinely
  not surfaced/validated after uninstall.
- LOW: phase-6 conformance leak regex broadened to catch absolute-home and
  Windows-backslash `.claude/(gsd-core|commands|agents|hooks)` paths, not only
  `~`/`$HOME` forward-slash forms.
- LOW: bug-2851 plan:post dispatch assertion now requires `--raw` (matched its
  stated contract).
- nit: plan-pre intel-step test duplicate assertion replaced with a distinct
  structured-output check.

Size baseline regenerated (execute-phase.md 93089 < 93166 frozen). Refs #1167, #1169.

Co-Authored-By: Claude Opus 4.8 <noreply@anthropic.com>

* test(#1169): make runtime-homes-descriptor-drive titles environment-independent

The descriptor-equivalence test embedded the absolute golden config path
(`os.homedir()`-derived) directly in each `test(...)` title, so titles differed
between macOS (`/Users/x/.claude`) and Docker (`/home/gsdtest/.claude`). Every
test PASSES on both platforms (15885/0 leaf tests each), but gsd-test-summary
compares results by title and reported 29+29 false "only in Mac / only in
Docker" discrepancies for tests that actually pass everywhere.

Move the golden path out of the title and into the assertion message (still
shown on failure); titles are now byte-identical across platforms so the
cross-platform comparator matches them. No assertion logic or golden values
changed. Refs #1169.

Co-Authored-By: Claude Opus 4.8 <noreply@anthropic.com>

* fix(#1169): derive TDD_MODE via gsd_run --active-cap, not node -e (fix prompt-injection CI gate)

The prior fix reverted execute-phase.md:181 from jq to `node -e` to close a
Codex HIGH (jq||echo-false silently disabling the MVP+TDD blocking gate on
jq-less runtimes) — but the CI prompt-injection scanner BLOCKS new `node -e` in
workflow markdown (inline code-exec = injection vector), turning the security
gate red. Both forms were wrong: node -e fails the scanner; jq fail-opens a
blocking safety gate; `config-get workflow.tdd_mode` is forbidden by the
conformance leak gate (tdd_mode is capability-owned).

Correct fix (what Codex recommended): a gsd_run-native boolean. Add an
`--active-cap <capId>` flag to `loop render-hooks <point>` that resolves hooks
the normal way and prints exactly `true`/`false` for whether a capId is active
— scanner-safe (canonical launcher, no inline code), node-reliable (no optional
jq to fail-open), and leak-free (render-hooks resolution, not config-get).
execute-phase.md:181 now `TDD_MODE=$(gsd_run loop render-hooks execute:post
--active-cap tdd)`. +5 behavioral tests for the flag.

Verified: prompt-injection-scan --diff origin/next → 0 findings; conformance
gate 13/13 (execute-phase.md 92934 < 93166); execute-mvp-tdd + tdd-mode +
loop-render-hooks 87/0; lint 0/0. Refs #1167, #1169.

Co-Authored-By: Claude Opus 4.8 <noreply@anthropic.com>

---------

Co-authored-by: Claude Opus 4.8 <noreply@anthropic.com>
2026-06-13 21:07:55 -04:00

544 lines
23 KiB
JavaScript

'use strict';
/**
* loop-hooks-verify-post-e2e.test.cjs
*
* E2E content tests for the verify:post hook point — ADR-857 phase 6.
*
* Coverage focus (backlog: hook-e2e-gaps.md § verify:post):
* - All-on: 3 hooks in registry order (nyquist → security → ui) with
* correct kind/ref.skill/onError (halt for nyquist+security, skip for ui)
* - No-config: schema defaults activate all 3
* - Per-key false: each of the 3 BVA cases excludes only that one step
* - All-false: empty activeHooks + valid envelope shape
* - Surface-disable (via capabilityStatesById on pure resolver): ui/security
* cluster excluded; remaining steps correct
* - Malformed config.json: falls back to schema defaults (3 active)
* - Deterministic ordering: two calls produce identical activeHooks arrays
*
* Hard rules enforced here:
* - Every test drives real resolver or CLI subprocess — no readFileSync source-grep
* - Genuine assertions: negative/BVA cases assert the SPECIFIC differing value
* - Each test owns its own fixture (isolated tmpDir); cleanup in afterEach
*/
const { describe, test, before, after, afterEach } = require('node:test');
const { cleanup } = require('./helpers.cjs');
const assert = require('node:assert/strict');
const fs = require('node:fs');
const os = require('node:os');
const path = require('node:path');
const { spawnSync } = require('node:child_process');
// ── Real modules under test ────────────────────────────────────────────────────
const {
resolveLoopHooks,
renderLoopHooks,
} = require('../gsd-core/bin/lib/loop-resolver.cjs');
const realRegistry = require('../gsd-core/bin/lib/capability-registry.cjs');
// ── CLI path ───────────────────────────────────────────────────────────────────
const GSD_TOOLS = path.join(__dirname, '..', 'gsd-core', 'bin', 'gsd-tools.cjs');
// ── Env hermeticity (strip ambient GSD_ vars that skew planning dir lookups) ──
const CLEAN_ENV = Object.fromEntries(
Object.entries(process.env).filter(([k]) => !k.startsWith('GSD_')),
);
/**
* Invoke gsd-tools CLI with spawnSync and return the parsed result.
* Always use CLEAN_ENV to avoid ambient GSD_ env vars redirecting planning paths.
*/
function runCli(args, cwd) {
const result = spawnSync(process.execPath, [GSD_TOOLS, ...args], {
cwd,
encoding: 'utf8',
env: CLEAN_ENV,
timeout: 60000,
});
return result;
}
/** Create a temp dir with a .planning/ subdirectory (no config.json). */
function makeTmpProject() {
const d = fs.mkdtempSync(path.join(os.tmpdir(), 'vpost-e2e-'));
fs.mkdirSync(path.join(d, '.planning'), { recursive: true });
return d;
}
/** Write .planning/config.json with the given object. */
function writeConfig(tmpDir, cfg) {
fs.writeFileSync(
path.join(tmpDir, '.planning', 'config.json'),
JSON.stringify(cfg),
'utf8',
);
}
// ── Fixtures shared across all-on and ordering tests ─────────────────────────
let allOnDir; // .planning/config.json with all three verify:post flags = true
let noConfigDir; // .planning/ but NO config.json
let allOffDir; // all three flags explicitly false
before(() => {
allOnDir = makeTmpProject();
writeConfig(allOnDir, {
workflow: { nyquist_validation: true, security_enforcement: true, ui_review: true },
});
noConfigDir = makeTmpProject();
// No config.json — schema defaults (all true) should activate all three
allOffDir = makeTmpProject();
writeConfig(allOffDir, {
workflow: { nyquist_validation: false, security_enforcement: false, ui_review: false },
});
});
after(() => {
for (const d of [allOnDir, noConfigDir, allOffDir]) {
if (d) cleanup(d);
}
});
// Per-test isolation: each test creates its own dir; afterEach cleans it up.
let perTestDir = null;
afterEach(() => {
if (perTestDir) {
cleanup(perTestDir);
perTestDir = null;
}
});
// ─── 1. All-on: three hooks in correct order with full typed shape ─────────────
describe('verify:post — all-on config activates all three steps in registry order', () => {
test('[happy] CLI returns 3 active hooks: nyquist→security→ui with correct capId, kind, ref.skill', () => {
const result = runCli(
['loop', 'render-hooks', 'verify:post', '--raw', '--cwd', allOnDir],
allOnDir,
);
assert.strictEqual(result.status, 0, `CLI exited non-zero: ${result.stderr}`);
const envelope = JSON.parse(result.stdout.trim());
assert.strictEqual(envelope.point, 'verify:post');
assert.strictEqual(envelope.activeHooks.length, 3,
`Expected 3 active hooks, got ${envelope.activeHooks.length}: ${JSON.stringify(envelope.activeHooks.map(h => h.capId))}`);
// Step 1: nyquist
const [nyquist, security, ui] = envelope.activeHooks;
assert.strictEqual(nyquist.capId, 'nyquist');
assert.strictEqual(nyquist.kind, 'step');
assert.strictEqual(nyquist.ref.skill, 'validate-phase');
assert.strictEqual(nyquist.onError, 'halt');
// Step 2: security
assert.strictEqual(security.capId, 'security');
assert.strictEqual(security.kind, 'step');
assert.strictEqual(security.ref.skill, 'secure-phase');
assert.strictEqual(security.onError, 'halt');
// Step 3: ui
assert.strictEqual(ui.capId, 'ui');
assert.strictEqual(ui.kind, 'step');
assert.strictEqual(ui.ref.skill, 'ui-review');
assert.strictEqual(ui.onError, 'skip');
});
test('[happy] CLI returns rendered markdown with Step 1/2/3 in correct order', () => {
const result = runCli(
['loop', 'render-hooks', 'verify:post', '--raw', '--cwd', allOnDir],
allOnDir,
);
assert.strictEqual(result.status, 0);
const envelope = JSON.parse(result.stdout.trim());
// Rendered text must contain all three steps in correct order
const { rendered } = envelope;
assert.ok(typeof rendered === 'string' && rendered.length > 0, 'rendered must be non-empty string');
const step1Pos = rendered.indexOf('validate-phase');
const step2Pos = rendered.indexOf('secure-phase');
const step3Pos = rendered.indexOf('ui-review');
assert.ok(step1Pos < step2Pos, `nyquist (pos ${step1Pos}) must come before security (pos ${step2Pos}) in rendered`);
assert.ok(step2Pos < step3Pos, `security (pos ${step2Pos}) must come before ui (pos ${step3Pos}) in rendered`);
// Rendered must NOT be the placeholder (all hooks active)
assert.ok(
!rendered.includes('_No active hooks at verify:post._'),
'rendered must not be the empty-hooks placeholder when all are active',
);
});
});
// ─── 2. No-config: schema defaults activate all 3 ─────────────────────────────
describe('verify:post — no config.json falls back to schema defaults (all three active)', () => {
test('[happy] CLI with no config.json returns 3 active hooks via schema default=true', () => {
const result = runCli(
['loop', 'render-hooks', 'verify:post', '--raw', '--cwd', noConfigDir],
noConfigDir,
);
assert.strictEqual(result.status, 0, `CLI exited non-zero: ${result.stderr}`);
const envelope = JSON.parse(result.stdout.trim());
assert.strictEqual(envelope.point, 'verify:post');
assert.strictEqual(envelope.activeHooks.length, 3,
`Schema defaults should activate 3 hooks, got ${envelope.activeHooks.length}`);
// Verify capIds — schema default=true for all three
const capIds = envelope.activeHooks.map(h => h.capId);
assert.deepEqual(capIds, ['nyquist', 'security', 'ui'],
`Expected ['nyquist','security','ui'], got ${JSON.stringify(capIds)}`);
});
test('[happy] pure resolveLoopHooks with realRegistry and empty config activates all 3 (schema default path)', () => {
const resolved = resolveLoopHooks({
point: 'verify:post',
registry: realRegistry,
config: {},
});
assert.strictEqual(resolved.point, 'verify:post');
assert.strictEqual(resolved.activeHooks.length, 3,
`Expected 3 active hooks via schema default, got ${resolved.activeHooks.length}`);
assert.deepEqual(
resolved.activeHooks.map(h => h.capId),
['nyquist', 'security', 'ui'],
);
});
});
// ─── 3. All-false: empty hooks + valid 3-key envelope ─────────────────────────
describe('verify:post — all three flags explicitly false returns empty activeHooks', () => {
test('[negative] CLI with all-false config returns activeHooks:[] and placeholder rendered', () => {
const result = runCli(
['loop', 'render-hooks', 'verify:post', '--raw', '--cwd', allOffDir],
allOffDir,
);
assert.strictEqual(result.status, 0, `CLI exited non-zero: ${result.stderr}`);
const envelope = JSON.parse(result.stdout.trim());
// Genuine assertion: MUST be 0 (not 1 or 3) — verifies filtering actually works
assert.strictEqual(envelope.activeHooks.length, 0,
`Expected 0 hooks when all flags=false, got ${envelope.activeHooks.length}: ${JSON.stringify(envelope.activeHooks.map(h => h.capId))}`);
assert.deepEqual(envelope.activeHooks, []);
assert.strictEqual(envelope.rendered, '_No active hooks at verify:post._');
assert.strictEqual(envelope.point, 'verify:post');
});
test('[negative] pure resolveLoopHooks with all-false config returns empty activeHooks', () => {
const resolved = resolveLoopHooks({
point: 'verify:post',
registry: realRegistry,
config: { workflow: { nyquist_validation: false, security_enforcement: false, ui_review: false } },
});
// Must be exactly 0, not 1 or 3
assert.strictEqual(resolved.activeHooks.length, 0);
assert.strictEqual(renderLoopHooks(resolved), '_No active hooks at verify:post._');
});
});
// ─── 4. BVA: per-key false excludes only that one step ────────────────────────
describe('verify:post — per-key BVA: each false excludes only that single step', () => {
test('[bva] nyquist_validation=false excludes ONLY nyquist; security+ui remain (length=2)', () => {
perTestDir = makeTmpProject();
writeConfig(perTestDir, {
workflow: { nyquist_validation: false, security_enforcement: true, ui_review: true },
});
const result = runCli(
['loop', 'render-hooks', 'verify:post', '--raw', '--cwd', perTestDir],
perTestDir,
);
assert.strictEqual(result.status, 0);
const envelope = JSON.parse(result.stdout.trim());
// Genuine BVA: must be exactly 2, not 3 or 0
assert.strictEqual(envelope.activeHooks.length, 2,
`Expected 2 hooks (security+ui), got ${envelope.activeHooks.length}: ${JSON.stringify(envelope.activeHooks.map(h => h.capId))}`);
const capIds = envelope.activeHooks.map(h => h.capId);
assert.ok(!capIds.includes('nyquist'), `nyquist must be absent when nyquist_validation=false, got ${JSON.stringify(capIds)}`);
assert.strictEqual(capIds[0], 'security', `First remaining hook must be security`);
assert.strictEqual(capIds[1], 'ui', `Second remaining hook must be ui`);
});
test('[bva] security_enforcement=false excludes ONLY security; nyquist+ui remain (length=2)', () => {
perTestDir = makeTmpProject();
writeConfig(perTestDir, {
workflow: { nyquist_validation: true, security_enforcement: false, ui_review: true },
});
const result = runCli(
['loop', 'render-hooks', 'verify:post', '--raw', '--cwd', perTestDir],
perTestDir,
);
assert.strictEqual(result.status, 0);
const envelope = JSON.parse(result.stdout.trim());
// Genuine BVA: must be exactly 2, not 3 or 0
assert.strictEqual(envelope.activeHooks.length, 2,
`Expected 2 hooks (nyquist+ui), got ${envelope.activeHooks.length}: ${JSON.stringify(envelope.activeHooks.map(h => h.capId))}`);
const capIds = envelope.activeHooks.map(h => h.capId);
assert.ok(!capIds.includes('security'), `security must be absent when security_enforcement=false, got ${JSON.stringify(capIds)}`);
assert.strictEqual(capIds[0], 'nyquist', `First remaining hook must be nyquist`);
assert.strictEqual(capIds[1], 'ui', `Second remaining hook must be ui`);
});
test('[bva] ui_review=false excludes ONLY ui; nyquist+security remain (length=2)', () => {
perTestDir = makeTmpProject();
writeConfig(perTestDir, {
workflow: { nyquist_validation: true, security_enforcement: true, ui_review: false },
});
const result = runCli(
['loop', 'render-hooks', 'verify:post', '--raw', '--cwd', perTestDir],
perTestDir,
);
assert.strictEqual(result.status, 0);
const envelope = JSON.parse(result.stdout.trim());
// Genuine BVA: must be exactly 2, not 3 or 0
assert.strictEqual(envelope.activeHooks.length, 2,
`Expected 2 hooks (nyquist+security), got ${envelope.activeHooks.length}: ${JSON.stringify(envelope.activeHooks.map(h => h.capId))}`);
const capIds = envelope.activeHooks.map(h => h.capId);
assert.ok(!capIds.includes('ui'), `ui must be absent when ui_review=false, got ${JSON.stringify(capIds)}`);
assert.strictEqual(capIds[0], 'nyquist', `First remaining hook must be nyquist`);
assert.strictEqual(capIds[1], 'security', `Second remaining hook must be security`);
});
});
// ─── 5. Surface-disable via capabilityStatesById (pure resolver) ──────────────
describe('verify:post — surface-disable: capabilityStatesById filters hooks', () => {
test('[negative] ui disabled via capabilityStatesById→enabled:false excludes ui step; nyquist+security remain', () => {
const capabilityStatesById = new Map([
['nyquist', { enabled: true }],
['security', { enabled: true }],
['ui', { enabled: false }],
]);
const resolved = resolveLoopHooks({
point: 'verify:post',
registry: realRegistry,
config: { workflow: { nyquist_validation: true, security_enforcement: true, ui_review: true } },
capabilityStatesById,
});
// Genuine assertion: must be 2 (not 3) — proves surface filter excludes ui
assert.strictEqual(resolved.activeHooks.length, 2,
`Expected 2 hooks with ui disabled, got ${resolved.activeHooks.length}: ${JSON.stringify(resolved.activeHooks.map(h => h.capId))}`);
const capIds = resolved.activeHooks.map(h => h.capId);
assert.ok(!capIds.includes('ui'), `ui must be filtered out when capability disabled`);
assert.strictEqual(capIds[0], 'nyquist');
assert.strictEqual(capIds[1], 'security');
});
test('[negative] security disabled via capabilityStatesById excludes security step; nyquist+ui remain', () => {
const capabilityStatesById = new Map([
['nyquist', { enabled: true }],
['security', { enabled: false }],
['ui', { enabled: true }],
]);
const resolved = resolveLoopHooks({
point: 'verify:post',
registry: realRegistry,
config: { workflow: { nyquist_validation: true, security_enforcement: true, ui_review: true } },
capabilityStatesById,
});
// Genuine: must be 2 (not 3) — proves security cluster exclusion
assert.strictEqual(resolved.activeHooks.length, 2,
`Expected 2 hooks with security disabled, got ${resolved.activeHooks.length}`);
const capIds = resolved.activeHooks.map(h => h.capId);
assert.ok(!capIds.includes('security'), `security must be filtered out when capability disabled`);
assert.strictEqual(capIds[0], 'nyquist');
assert.strictEqual(capIds[1], 'ui');
});
test('[empty-resolution] all three disabled via capabilityStatesById returns empty activeHooks with valid envelope', () => {
const capabilityStatesById = new Map([
['nyquist', { enabled: false }],
['security', { enabled: false }],
['ui', { enabled: false }],
]);
const resolved = resolveLoopHooks({
point: 'verify:post',
registry: realRegistry,
config: { workflow: { nyquist_validation: true, security_enforcement: true, ui_review: true } },
capabilityStatesById,
});
assert.strictEqual(resolved.point, 'verify:post');
assert.deepEqual(resolved.activeHooks, []);
assert.strictEqual(renderLoopHooks(resolved), '_No active hooks at verify:post._');
});
});
// ─── 6. Malformed config.json: falls back to schema defaults ──────────────────
describe('verify:post — malformed config.json: schema defaults fire (3 active, no crash)', () => {
test('[negative] CLI with malformed config.json exits 0 and returns all 3 hooks via schema defaults', () => {
perTestDir = makeTmpProject();
fs.writeFileSync(
path.join(perTestDir, '.planning', 'config.json'),
'{ broken json',
'utf8',
);
const result = runCli(
['loop', 'render-hooks', 'verify:post', '--raw', '--cwd', perTestDir],
perTestDir,
);
assert.strictEqual(result.status, 0, `CLI must not crash on malformed config: ${result.stderr}`);
const envelope = JSON.parse(result.stdout.trim());
assert.strictEqual(envelope.point, 'verify:post');
// Schema defaults (all true) must activate all 3 when config.json parse fails
assert.strictEqual(envelope.activeHooks.length, 3,
`Expected 3 hooks via schema defaults on malformed config, got ${envelope.activeHooks.length}`);
const capIds = envelope.activeHooks.map(h => h.capId);
assert.deepEqual(capIds, ['nyquist', 'security', 'ui']);
});
});
// ─── 7. Deterministic ordering: two calls produce identical results ────────────
describe('verify:post — deterministic ordering: repeated calls produce identical activeHooks', () => {
test('[happy] two resolveLoopHooks calls return identical activeHooks arrays (order stability)', () => {
const config = {
workflow: { nyquist_validation: true, security_enforcement: true, ui_review: true },
};
const first = resolveLoopHooks({ point: 'verify:post', registry: realRegistry, config });
const second = resolveLoopHooks({ point: 'verify:post', registry: realRegistry, config });
// Genuine: both must have exactly the same structure
assert.deepEqual(first.activeHooks, second.activeHooks,
'Two resolver calls must produce identical activeHooks (determinism)');
assert.deepEqual(
first.activeHooks.map(h => h.capId),
['nyquist', 'security', 'ui'],
'Order must be nyquist→security→ui',
);
});
test('[happy] two CLI invocations return identical stdout (CLI-level determinism)', () => {
const call1 = runCli(
['loop', 'render-hooks', 'verify:post', '--raw', '--cwd', allOnDir],
allOnDir,
);
const call2 = runCli(
['loop', 'render-hooks', 'verify:post', '--raw', '--cwd', allOnDir],
allOnDir,
);
assert.strictEqual(call1.status, 0);
assert.strictEqual(call2.status, 0);
const env1 = JSON.parse(call1.stdout.trim());
const env2 = JSON.parse(call2.stdout.trim());
assert.deepEqual(env1.activeHooks, env2.activeHooks,
'Two CLI calls must produce identical activeHooks');
assert.strictEqual(env1.rendered, env2.rendered,
'Two CLI calls must produce identical rendered output');
});
});
// ─── 8. onError fields per-hook (halt for nyquist+security, skip for ui) ──────
describe('verify:post — onError semantics: halt for nyquist+security, skip for ui', () => {
test('[bva] onError is exactly "halt" for nyquist, "halt" for security, "skip" for ui — pure resolver', () => {
const resolved = resolveLoopHooks({
point: 'verify:post',
registry: realRegistry,
config: { workflow: { nyquist_validation: true, security_enforcement: true, ui_review: true } },
});
assert.strictEqual(resolved.activeHooks.length, 3);
// Genuine BVA: each onError must match the exact canonical value
assert.strictEqual(resolved.activeHooks[0].onError, 'halt',
`nyquist onError must be 'halt', got '${resolved.activeHooks[0].onError}'`);
assert.strictEqual(resolved.activeHooks[1].onError, 'halt',
`security onError must be 'halt', got '${resolved.activeHooks[1].onError}'`);
assert.strictEqual(resolved.activeHooks[2].onError, 'skip',
`ui onError must be 'skip', got '${resolved.activeHooks[2].onError}'`);
});
test('[bva] CLI envelope preserves onError values in the correct field position', () => {
const result = runCli(
['loop', 'render-hooks', 'verify:post', '--raw', '--cwd', allOnDir],
allOnDir,
);
assert.strictEqual(result.status, 0);
const envelope = JSON.parse(result.stdout.trim());
// Genuine BVA: assert specific onError value at each position, not just presence
assert.strictEqual(envelope.activeHooks[0].onError, 'halt');
assert.strictEqual(envelope.activeHooks[1].onError, 'halt');
assert.strictEqual(envelope.activeHooks[2].onError, 'skip');
});
});
// ─── 9. Envelope shape: exactly 3 keys, no spurious 'warnings' ────────────────
describe('verify:post — envelope shape pins Hyrum\'s Law contract', () => {
test('[happy] all-on CLI response has exactly 3 envelope keys: point, activeHooks, rendered', () => {
const result = runCli(
['loop', 'render-hooks', 'verify:post', '--raw', '--cwd', allOnDir],
allOnDir,
);
assert.strictEqual(result.status, 0);
const envelope = JSON.parse(result.stdout.trim());
// When state.warnings is empty, the envelope must have exactly 3 keys
const keys = Object.keys(envelope).sort();
assert.deepEqual(keys, ['activeHooks', 'point', 'rendered'],
`Envelope must have exactly 3 keys, got: ${JSON.stringify(keys)}`);
});
test('[negative] all-off CLI response envelope still has exactly 3 keys (no extra warnings key)', () => {
const result = runCli(
['loop', 'render-hooks', 'verify:post', '--raw', '--cwd', allOffDir],
allOffDir,
);
assert.strictEqual(result.status, 0);
const envelope = JSON.parse(result.stdout.trim());
// All-false path: 3 keys, not more
const keys = Object.keys(envelope).sort();
assert.deepEqual(keys, ['activeHooks', 'point', 'rendered'],
`Empty-hooks envelope must have exactly 3 keys, got: ${JSON.stringify(keys)}`);
assert.strictEqual(envelope.point, 'verify:post');
assert.deepEqual(envelope.activeHooks, []);
});
});
// ─── 10. Real registry byLoopPoint shape check (no drift guard) ───────────────
describe('verify:post — real registry has exactly 3 steps and 0 contributions+gates', () => {
test('[happy] realRegistry.byLoopPoint[verify:post] has 3 steps, 0 contributions, 0 gates', () => {
const entry = realRegistry.byLoopPoint['verify:post'];
assert.ok(entry, 'verify:post must exist in registry');
assert.strictEqual(entry.steps.length, 3,
`Expected 3 steps at verify:post, got ${entry.steps.length}`);
assert.strictEqual(entry.contributions.length, 0,
`Expected 0 contributions at verify:post, got ${entry.contributions.length}`);
assert.strictEqual(entry.gates.length, 0,
`Expected 0 gates at verify:post, got ${entry.gates.length}`);
});
test('[happy] registry steps at verify:post have correct capIds in order', () => {
const entry = realRegistry.byLoopPoint['verify:post'];
const capIds = entry.steps.map(s => s.capId);
assert.deepEqual(capIds, ['nyquist', 'security', 'ui'],
`Registry must have steps in nyquist→security→ui order, got ${JSON.stringify(capIds)}`);
});
});