* test(#1168): make phase-6 gate un-gameable — reject empty stubs + require loop shrink The migration assertion previously checked only role==feature, so a registration-only stub (empty hooks, logic left inline) would turn the gate green while phase 6 stayed incomplete — the exact false-completion pattern this gate exists to prevent. Strengthen it: each ADR-named feature must OWN its behavior (>=1 hook, or a command family); and plan-phase.md/execute-phase.md must shrink strictly below their frozen pre-phase-6 sizes (94519/93166 LF bytes), which also defeats double-run gaming (declare a hook but keep the inline block -> file does not shrink -> red). Gate now 5 pass / 4 fail (orphaned execute:wave:post, empty/unregistered features, config-key leaks, no shrink). Green is now reachable only by REAL migration. Refs #1168, #1169. Co-Authored-By: Claude Opus 4.8 <noreply@anthropic.com> * feat(#1169): migrate gap-analysis to a Capability (plan:post gate) First real ADR-857 phase-6 migration (pattern-defining tracer). gap-analysis moves from an inline post_planning_gaps branch in plan-phase.md to a real plan:post gate Capability: - capabilities/gap-analysis/capability.json: role:feature, plan:post gate (when=workflow.post_planning_gaps, blocking:false advisory), OWNS workflow.post_planning_gaps (federated out of central schema). - plan-phase.md: inline config-get + gsd_run gap-analysis block replaced with a plan:post render-hooks call site dispatching the gate; file shrinks 94519->93279. - src/check-command-router.cts: cmdGapAnalysisPlanPost runs the real gap analysis via gap-checker. - post_planning_gaps removed from central manifest; resolves via federated config (default true preserved). - tests/post-planning-gaps-2493: re-pointed to assert capability ownership. Verified: gate 5 pass / 4 fail (gap-analysis cleared from migration, plan:post-orphan, config-leak, and plan-phase shrink checks); loadConfig still returns post_planning_gaps=true; check command runs real analysis; 392/392 in the config/registry/federation/router net. Refs #1169. Co-Authored-By: Claude Opus 4.8 <noreply@anthropic.com> * feat(#1169): migrate profile-pipeline to a command-family Capability ADR-857 Decision 7: profile-pipeline becomes a command-family Capability (like audit/intel/graphify). capabilities/profile-pipeline/capability.json declares an 8-command family (scan-sessions, extract-messages, profile-sample, write-profile, profile-questionnaire, generate-dev-preferences, generate-claude-profile, generate-claude-md) backed by a new gsd-core/bin/lib/profile-pipeline-command-router.cjs; the inline case arms are removed from gsd-tools.cjs. Owns profile-pipeline.enabled (federated). Verified: registry shows role:feature with commands.length=8; scan-sessions/profile-sample run live via the family; gate cleared profile-pipeline from the empty-stub failure (only tdd/schema-gate/drift remain); 296/296 registry+inventory+gsd-tools tests; lint 0 errors. Refs #1169. Co-Authored-By: Claude Opus 4.8 <noreply@anthropic.com> * fix(#1167): wire execute:wave:post + implement ui.safety-gate check Revives the second dead gate from #1167: ui.gates@execute:wave:post was declared but never dispatched AND its check.query (ui.safety-gate) was unimplemented. Adds the per-wave execute:wave:post render-hooks call site in execute-phase.md (fires after each wave's merge/cleanup, before the next forks) and implements cmdUiSafetyGate (frontend + UI-SPEC aware, mirrors cmdUiPlanGate) in check-command-router. +17 regression tests. Verified: phase-6 orphaned-points conformance test now PASSES (gate 6 pass / 3 fail); ui-safety-gate routable in dot+hyphen forms; check-ui-safety-gate 17/17, check-ui-plan-gate 18/18; lint 0 errors. Refs #1167, #1168. Co-Authored-By: Claude Opus 4.8 <noreply@anthropic.com> * feat(#1169): migrate drift (schema + codebase) to execute:wave:post gates Removes the inline schema_drift_gate + codebase_drift_gate steps (77 lines) from execute-phase.md; drift becomes a Capability with two execute:wave:post gates (verify.schema-drift blocking, verify.codebase-drift advisory) dispatched via the per-wave render-hooks call site. check-command-router routes verify.schema-drift / verify.codebase-drift to the real detectors. Federates workflow.drift_threshold / drift_action / schema_drift_gate out of central. Also fixes the execute:wave:post dispatch prose to run NON-blocking (advisory) gates too — the prior version only ran blocking gates, which would have silently dropped the codebase-drift advisory after its inline step was removed. Behavior preserved. Verified: gate 7 pass / 2 fail (drift cleared from stub + config-leak; execute-phase.md 92297 < 93166 frozen -> shrink passes); both drift checks run real detection; loadConfig defaults preserved (threshold=3, action=warn, gate=true); drift-detection 56/56 + schema-drift 34/34; lint 0 errors. Refs #1169. Co-Authored-By: Claude Opus 4.8 <noreply@anthropic.com> * feat(#1169): migrate tdd to a Capability (plan:pre contribution + execute:post gate) tdd becomes a real Capability: a plan:pre contribution injects the <tdd_mode_active> planner guidance (rendered from PLAN_PRE_HOOKS_JSON like security's contribution), and an execute:post gate (tdd.review-checkpoint, advisory) runs the real end-of-phase RED/GREEN review via a new check-command handler. Inline tdd_mode reads + the inline planner block + the tdd_review_checkpoint step are removed; workflow.tdd_mode is federated out of central. The MVP+TDD per-task RED-commit gate is preserved — TDD_MODE is now derived from the execute:post hooks (capId==tdd active), not an inline config-get. BEHAVIOR CHANGE (documented, not silent): the --tdd CLI flag now persists workflow.tdd_mode=true via config-set instead of being per-invocation. Rationale: tdd is now a config-toggled Capability, and env vars do not persist across the workflow's separate bash blocks (config does), so an ephemeral override isn't cleanly achievable; --tdd therefore enables the tdd capability, consistent with how all capabilities are toggled. Verified: gate 7 pass / 2 fail (tdd cleared from stub + config-leak; plan-phase + execute-phase both < frozen sizes); contribution injection + execute:post gate dispatch wired; MVP+TDD gate preserved; tdd.review-checkpoint runs real review; full unit suite 556/0; lint 0 errors. Refs #1169. Co-Authored-By: Claude Opus 4.8 <noreply@anthropic.com> * feat(#1169): migrate schema-gate to a plan:pre contribution Capability The plan-time schema-push detection (former plan-phase.md §5.7) becomes a schema-gate Capability: a plan:pre contribution (into:planner, when:workflow.schema_push_detection) whose fragment carries the full ORM-detection + [BLOCKING] schema-push-task injection logic, rendered into the planner via the existing plan:pre render-hooks dispatch. The inline §5.7 block is removed (plan-phase.md 94519->90445). workflow.schema_push_detection is a new capability-owned (federated) key, default true. (The execute-side schema-drift gate was migrated separately into the drift capability.) Verified: registry inlines the fragment (len 2704) so it is actually delivered at plan:pre; gate 8 pass / 1 fail — all 5 ADR-named features now real Capabilities, only the config-leak test remains (intel/security, next unit). Refs #1169. Co-Authored-By: Claude Opus 4.8 <noreply@anthropic.com> * feat(#1169): close the 3 capability config-key leaks — phase-6 gate now GREEN Removes the last inline config-get reads of capability-owned keys from plan-phase.md. security_asvs_level/security_block_on now flow through the security plan:pre contribution via a new loop-resolver configValues mechanism (resolves declared config keys with the same 4-level precedence as activation and attaches them to the rendered hook); the §5.55 banner reads them from PLAN_PRE_HOOKS_JSON. intel.enabled becomes a real intel plan:pre step (ref.command: intel api-surface) dispatched via render-hooks; the inline intel branch is gone. gen-capability-registry now validates ref.command as a third dispatch shape. Verified: phase-6 capstone conformance gate is FULLY GREEN (9/0); 3 leaks gone (grep=0); security configValues resolve to {2,medium}/default {1,high}; intel step present only when enabled; loop-render-hooks 62/0, capability-registry 287/0, capability-state/federated-config 113/0; lint 0 errors. Closes the migration half of #1169. Refs #1139, #1167, #1168. Co-Authored-By: Claude Opus 4.8 <noreply@anthropic.com> * fix(#1169): address adversarial review — restore schema-drift block, generic planner injection, uniform gate contract Adversarial review caught 2 real regressions the green gate missed: (1) schema-drift no longer blocked — the execute:wave:post dispatch read GATE_RESULT.block but verify.schema-drift emitted drift_detected/blocking, and onError:skip wrongly bypassed positive blocks; (2) only tdd's plan:pre contribution was injected into the planner, dropping schema-gate's schema-push detection and security's threat-model guidance. Fixes: (A) every gate check returns a uniform boolean 'block' under --raw (the dispatch form), with advisory gates (tdd/gap) carrying their report in 'message'; (B) gate-dispatch contract corrected at all sites — onError governs command errors only, a blocking gate's positive block always halts; (C) generic planner injection of all plan:pre contributions where into=='planner' (tdd + schema-gate + security incl configValues); (D) two new conformance assertions: planner contributions injected generically + every gate check.query returns boolean block under --raw. Verified: gate 11/11; all 6 gate checks return boolean block under --raw; full suite 595/0; lint 0 errors. Refs #1167, #1168, #1169. Co-Authored-By: Claude Opus 4.8 <noreply@anthropic.com> * fix(#1169): restore MVP+TDD end-of-phase blocking escalation (2nd adversarial pass) The migrated tdd execute:post gate is statically blocking:false, but the contract (references/execute-mvp-tdd.md + CONTEXT.md) requires the end-of-phase TDD review to ESCALATE from advisory to blocking when MVP_MODE && TDD_MODE && a TDD plan misses a RED/GREEN commit. The migration prose had downgraded this to a 'strong advisory recommendation' — silent loss of the blocking escalation. Restore it: the tdd-gate dispatch now refuses to mark the phase complete (Phase blocked message) under MVP+TDD when GATE_RESULT.block is true; advisory otherwise. Also strengthen tests/execute-mvp-tdd-gate.test.cjs: hasBlockingEscalation previously matched any line with 'blocking'+'mvp+tdd' (so 'advisory (blocking: false) ... under MVP+TDD' was a false green); now it requires the real refusal semantics ('refuse to mark the phase complete' / 'phase blocked'). Caught by 2nd adversarial review pass. Verified: execute-phase.md 92702 < 93166 frozen; mvp-tdd-gate + phase-6 gate 19/0; full suite green; lint 0 errors. Refs #1169. Co-Authored-By: Claude Opus 4.8 <noreply@anthropic.com> * fix(#1169): restore MVP+TDD proceed-block, codebase auto-remap, schema skip-flag (3rd adversarial pass) 3rd adversarial pass found 4 more silent regressions: (1) the tdd MVP+TDD 'refuse to mark complete' was nullified by a downstream 'ALWAYS proceed regardless of gate results' line — proceed is now conditional (stops on an active MVP+TDD block); (2) the test now asserts the proceed is NOT an unconditional override; (3) codebase-drift auto-remap (spawn gsd-codebase-mapper when drift_action=auto-remap) was dropped — the execute:wave:post advisory dispatch now consumes spawn_mapper/directive; (4) GSD_SKIP_SCHEMA_CHECK bypass was lost from the gate path — cmdVerifySchemaDrift now honors the env var (block:false when set). Verified: no unconditional proceed; GSD_SKIP_SCHEMA_CHECK=true -> block:false; gate 11/11 + mvp-tdd 9/9; full suite 569/0; lint 0; execute-phase.md 93109 < 93166. Refs #1169. Co-Authored-By: Claude Opus 4.8 <noreply@anthropic.com> * fix(#1169): init.cts reads federated config keys from nested path (4th adversarial pass) Config federation moved tdd_mode/research/nyquist_validation from flat config.<key> to nested config.workflow.<key>, but src/init.cts still read them flat — so init.plan-phase/init.execute-phase emitted tdd_mode:false / research_enabled:undefined / nyquist:undefined regardless of config (a public command-contract regression; the migrated loops use render-hooks so enforcement was unaffected). Read via config.workflow (type-safe Record cast). Now init reflects the same resolved values + federated defaults (research/nyquist default true) as the render-hooks path. Verified: build clean; init.plan-phase emits tdd_mode:true/research:false/nyquist:false for set config, defaults true for empty; full suite 591/0; lint 0. Refs #1169. Co-Authored-By: Claude Opus 4.8 <noreply@anthropic.com> * docs(#1169): add changeset for ADR-857 phase-6 completion (PR #1183) Co-Authored-By: Claude Opus 4.8 <noreply@anthropic.com> * fix(#1169): complete phase-6 migration fallout — restore TEXT_MODE, fix registry .claude leak, re-point stale workflow-contract tests The capability migration left real regressions and stale consumer tests that the per-module unit suite missed but the full cross-platform suite caught (27 failing tests): Real source regressions (fixed): - execute-phase.md lost its AskUserQuestion TEXT_MODE plain-text fallback when the inline schema_drift_gate step was removed — non-Claude runtimes would stall. Restored, and the execute:post gate-dispatch prose de-duplicated to cite the execute:wave:post contract (loop body shrinks below the frozen pre-phase-6 ceiling while keeping every onError/blocking nuance). - capabilities/tdd inline fragment hardcoded `@~/.claude/gsd-core/references/tdd.md`, baked verbatim into the committed capability-registry.cjs and leaked the install path on 11 non-Claude runtimes (registry .cjs is copied, not path-converted). Made the fragment path-free; regenerated the registry. The phase-6 conformance gate now guards this (no ~/.claude install path in any capability source or the generated registry). - plan-phase.md: removed a §5.7 stub re-added in error and routed Branch 2 to step 6 (schema-gate is a plan:pre capability, §5.7 is gone). Stale workflow-contract tests re-pointed to the capability dispatch they now must assert (behavior verified preserved in source first, assertions kept equal-or-stronger): bug-621 + bug-2851 (gap-analysis via gsd_run render-hooks plan:post + registry binding), feat-2527 (tdd_mode federated out of central), phase6-planning + plan-phase-ui-redirect (§5.6 bounded by ## 6.), plan-phase-drift-guard (intel when:intel.enabled skip branch). profile-pipeline-command-router.cjs un-ignored from eslint (hand-written, no TS source) + stale disable comments removed. Size baseline regenerated. Verified: full suite 15140 tests / 0 fail; lint 0 errors; conformance gate green legitimately. Refs #1139, #1167, #1168, #1169. Co-Authored-By: Claude Opus 4.8 <noreply@anthropic.com> * test(#1169): add ADR-857 E2E content-test coverage for the 12 loop points + capability deliverables Grounds the capability engine in behavioral E2E tests (drive the real render-hooks/check CLI + the real registry, assert typed result content — no source-grep), structured around what ADR-857 says to deliver. 207 tests; each genuineness-checked (flip the expectation, confirm it fails). Per-loop-point dispatch (7 files): empty-point negative-space across the 6 no-hook points; verify:post 3-step resolution+ordering+onError; plan:pre contribution/configValues + ui.plan-gate + intel; plan:post gap-analysis; execute:wave:post drift+ui gates via the check route (schema-drift block/skip, codebase-drift threshold BVA, auto-remap); execute:post tdd.review-checkpoint RED/GREEN; ship:pre security gate resolution + frontmatter-get predicate pieces. ADR-deliverable coverage (4 files): predicate boundary held (edge/prohibition probes stay core, not off-by-default Feature Capabilities — phase-6 exception); core loop runs with zero capabilities (all 12 points empty, init bundles resolve); contribution merge (multiple ordered <contribution from=> blocks); federated-config key removal on uninstall. federated-config allowlisted for its 3-file split (unit + integration + lifecycle). Refs #1139, #1167, #1168, #1169. Co-Authored-By: Claude Opus 4.8 <noreply@anthropic.com> * fix(#1169): remove dead drifted converter dups + address adversarial review Lint cleanup (root-caused, not waved off): src/runtime-artifact-conversion.cts carried 11 agent-converter functions (+5 orphaned consts/helpers) that were never exported, never called, and had silently DRIFTED from the live hand-authored copies in bin/install.js (one even referenced an undefined `claudeToCopilotTools`). Deleted the dead duplicates; install.js's live copies are untouched (it never imported these). Lint now 0 errors / 0 warnings. Adversarial-review (Codex) findings fixed: - HIGH: execute-phase.md TDD_MODE used `jq ... || echo false`, silently disabling the MVP+TDD blocking gate on jq-less runtimes. Reverted to the `node -e` form (node is guaranteed; matches the file's other node-e usages) so a missing optional tool can no longer fail-open a blocking safety path. - MEDIUM: federated-config-key-removal orphan-key test was vacuous (it skipped the orphan assertion). Now asserts the removed capability's key is genuinely not surfaced/validated after uninstall. - LOW: phase-6 conformance leak regex broadened to catch absolute-home and Windows-backslash `.claude/(gsd-core|commands|agents|hooks)` paths, not only `~`/`$HOME` forward-slash forms. - LOW: bug-2851 plan:post dispatch assertion now requires `--raw` (matched its stated contract). - nit: plan-pre intel-step test duplicate assertion replaced with a distinct structured-output check. Size baseline regenerated (execute-phase.md 93089 < 93166 frozen). Refs #1167, #1169. Co-Authored-By: Claude Opus 4.8 <noreply@anthropic.com> * test(#1169): make runtime-homes-descriptor-drive titles environment-independent The descriptor-equivalence test embedded the absolute golden config path (`os.homedir()`-derived) directly in each `test(...)` title, so titles differed between macOS (`/Users/x/.claude`) and Docker (`/home/gsdtest/.claude`). Every test PASSES on both platforms (15885/0 leaf tests each), but gsd-test-summary compares results by title and reported 29+29 false "only in Mac / only in Docker" discrepancies for tests that actually pass everywhere. Move the golden path out of the title and into the assertion message (still shown on failure); titles are now byte-identical across platforms so the cross-platform comparator matches them. No assertion logic or golden values changed. Refs #1169. Co-Authored-By: Claude Opus 4.8 <noreply@anthropic.com> * fix(#1169): derive TDD_MODE via gsd_run --active-cap, not node -e (fix prompt-injection CI gate) The prior fix reverted execute-phase.md:181 from jq to `node -e` to close a Codex HIGH (jq||echo-false silently disabling the MVP+TDD blocking gate on jq-less runtimes) — but the CI prompt-injection scanner BLOCKS new `node -e` in workflow markdown (inline code-exec = injection vector), turning the security gate red. Both forms were wrong: node -e fails the scanner; jq fail-opens a blocking safety gate; `config-get workflow.tdd_mode` is forbidden by the conformance leak gate (tdd_mode is capability-owned). Correct fix (what Codex recommended): a gsd_run-native boolean. Add an `--active-cap <capId>` flag to `loop render-hooks <point>` that resolves hooks the normal way and prints exactly `true`/`false` for whether a capId is active — scanner-safe (canonical launcher, no inline code), node-reliable (no optional jq to fail-open), and leak-free (render-hooks resolution, not config-get). execute-phase.md:181 now `TDD_MODE=$(gsd_run loop render-hooks execute:post --active-cap tdd)`. +5 behavioral tests for the flag. Verified: prompt-injection-scan --diff origin/next → 0 findings; conformance gate 13/13 (execute-phase.md 92934 < 93166); execute-mvp-tdd + tdd-mode + loop-render-hooks 87/0; lint 0/0. Refs #1167, #1169. Co-Authored-By: Claude Opus 4.8 <noreply@anthropic.com> --------- Co-authored-by: Claude Opus 4.8 <noreply@anthropic.com>
555 lines
21 KiB
JavaScript
555 lines
21 KiB
JavaScript
'use strict';
|
|
|
|
/**
|
|
* adr857-core-without-capabilities.test.cjs
|
|
*
|
|
* ADR-857 deliverable B — "the core loop ships and runs without any plug-in"
|
|
* (Consequences, §"Positive": "The core loop ships and runs without any plug-in;
|
|
* plan-phase.md/execute-phase.md shrink to the irreducible five steps.")
|
|
*
|
|
* Verified contracts:
|
|
* B1. All 12 canonical loop points return activeHooks:[] when every
|
|
* capability when-key is explicitly false (real registry, all-caps-off config).
|
|
* B2. The CLI `loop render-hooks <point>` exits 0 and emits activeHooks:[],
|
|
* placeholder rendered for representative points with all-caps-off config.
|
|
* B3. Init bundles for the 5-step loop's entry seam (plan-phase, execute-phase,
|
|
* verify-work) resolve with exit 0 and valid JSON when capabilities are off.
|
|
* B4. An EMPTY registry (byLoopPoint:{}) at all 12 points → activeHooks:[]
|
|
* (loop tolerates a capability-less install).
|
|
* B5. [BVA] Exactly one capability ON (tdd_mode) → that capability's points
|
|
* non-empty, all OTHER points still empty (caps are additive; core is baseline).
|
|
*
|
|
* RULESET: no readFileSync + .includes() on source files (source-grep ban).
|
|
* All assertions drive real exported functions / subprocess and inspect typed results.
|
|
*/
|
|
|
|
const { describe, test, before, after } = require('node:test');
|
|
const assert = require('node:assert/strict');
|
|
const fs = require('node:fs');
|
|
const os = require('node:os');
|
|
const path = require('node:path');
|
|
|
|
const { execFileSync } = require('child_process');
|
|
|
|
// ── Module under test ─────────────────────────────────────────────────────────
|
|
|
|
const {
|
|
resolveLoopHooks,
|
|
renderLoopHooks,
|
|
CANONICAL_POINTS,
|
|
} = require('../gsd-core/bin/lib/loop-resolver.cjs');
|
|
|
|
// Real registry (compiled from capabilities/ at build time)
|
|
const realRegistry = require('../gsd-core/bin/lib/capability-registry.cjs');
|
|
|
|
// ── Helpers from test harness ─────────────────────────────────────────────────
|
|
|
|
const { cleanup } = require('./helpers.cjs');
|
|
|
|
// ── Paths ─────────────────────────────────────────────────────────────────────
|
|
|
|
const GSD_TOOLS = path.join(__dirname, '..', 'gsd-core', 'bin', 'gsd-tools.cjs');
|
|
|
|
// ── Config fixtures ───────────────────────────────────────────────────────────
|
|
|
|
/**
|
|
* All 12 canonical loop points. Derived from the exported constant so the
|
|
* assertion set cannot drift from the resolver's own authoritative list.
|
|
*/
|
|
const ALL_12_POINTS = [...CANONICAL_POINTS];
|
|
|
|
/**
|
|
* All when-keys discovered from the real registry, set to false.
|
|
* Built by scanning every hook in every loop point's steps/contributions/gates arrays.
|
|
* This gives us a "caps-off" config that passes through the activation resolver as
|
|
* explicitly false rather than relying on missing-key default behaviour.
|
|
*
|
|
* Structure: nested (workflow.* → workflow:{...}, intel.enabled → intel:{enabled:false})
|
|
* because _getNestedConfigValue expects a nested object, not a flat dotted key.
|
|
*/
|
|
function buildAllFalseConfig() {
|
|
const workflow = {};
|
|
const intel = {};
|
|
for (const point of ALL_12_POINTS) {
|
|
const entry = realRegistry.byLoopPoint[point];
|
|
if (!entry) continue;
|
|
for (const kind of ['steps', 'contributions', 'gates']) {
|
|
for (const hook of entry[kind] || []) {
|
|
const when = hook.when;
|
|
if (typeof when !== 'string' || !when) continue;
|
|
if (when.startsWith('workflow.')) {
|
|
const key = when.slice('workflow.'.length);
|
|
workflow[key] = false;
|
|
} else if (when === 'intel.enabled') {
|
|
intel.enabled = false;
|
|
}
|
|
// Any future top-level keys would need extending here.
|
|
}
|
|
}
|
|
}
|
|
return { workflow, intel };
|
|
}
|
|
|
|
const ALL_FALSE_CONFIG = buildAllFalseConfig();
|
|
|
|
/**
|
|
* All-false config with tdd_mode: true.
|
|
* Only workflow.tdd_mode differs from ALL_FALSE_CONFIG.
|
|
*/
|
|
function buildTddOnlyConfig() {
|
|
return {
|
|
...ALL_FALSE_CONFIG,
|
|
workflow: { ...ALL_FALSE_CONFIG.workflow, tdd_mode: true },
|
|
};
|
|
}
|
|
|
|
// ── Helpers ───────────────────────────────────────────────────────────────────
|
|
|
|
/**
|
|
* Run gsd-tools subprocess and return { exitCode, output }.
|
|
* Does NOT throw on non-zero exit — let the test assert.
|
|
*/
|
|
function runCli(args, cwd) {
|
|
try {
|
|
const stdout = execFileSync(process.execPath, [GSD_TOOLS, ...args], {
|
|
cwd,
|
|
encoding: 'utf-8',
|
|
timeout: 30000,
|
|
});
|
|
return { exitCode: 0, output: stdout.trim() };
|
|
} catch (err) {
|
|
return {
|
|
exitCode: err.status ?? 1,
|
|
output: err.stdout?.toString().trim() ?? '',
|
|
error: err.stderr?.toString().trim() ?? '',
|
|
};
|
|
}
|
|
}
|
|
|
|
/** Create a temp project dir with a .planning/ sub-dir. */
|
|
function makeProject(configJson = null) {
|
|
const dir = fs.mkdtempSync(path.join(os.tmpdir(), 'adr857-b-'));
|
|
const planning = path.join(dir, '.planning');
|
|
fs.mkdirSync(planning, { recursive: true });
|
|
if (configJson !== null) {
|
|
fs.writeFileSync(path.join(planning, 'config.json'), JSON.stringify(configJson), 'utf8');
|
|
}
|
|
return dir;
|
|
}
|
|
|
|
/** Remove a temp dir safely. */
|
|
function removeTmp(dir) {
|
|
if (dir) cleanup(dir);
|
|
}
|
|
|
|
// ─────────────────────────────────────────────────────────────────────────────
|
|
// B1. [happy/aggregate] All 12 points → activeHooks:[] with all-caps-off config
|
|
// ─────────────────────────────────────────────────────────────────────────────
|
|
|
|
describe('B1 — real registry, all-caps-off config: every canonical point resolves to activeHooks:[]', () => {
|
|
test('all 12 CANONICAL_POINTS return activeHooks:[] simultaneously when every capability when-key is false', () => {
|
|
const failures = [];
|
|
for (const point of ALL_12_POINTS) {
|
|
const result = resolveLoopHooks({
|
|
point,
|
|
registry: realRegistry,
|
|
config: ALL_FALSE_CONFIG,
|
|
});
|
|
|
|
// Shape guard — result must be an object with an array
|
|
assert.ok(result && typeof result === 'object', `${point}: result must be an object`);
|
|
assert.ok(Array.isArray(result.activeHooks), `${point}: activeHooks must be an array`);
|
|
|
|
if (result.activeHooks.length !== 0) {
|
|
failures.push({
|
|
point,
|
|
count: result.activeHooks.length,
|
|
capIds: result.activeHooks.map(h => h.capId),
|
|
});
|
|
}
|
|
}
|
|
|
|
// Genuine assertion: if any point has activeHooks, report them concretely.
|
|
// This fails on regression to the specific wrong value, not just "not empty".
|
|
assert.deepStrictEqual(
|
|
failures,
|
|
[],
|
|
`Expected zero active hooks at all 12 points with all-caps-off config but got: ${JSON.stringify(failures)}`,
|
|
);
|
|
});
|
|
|
|
test('CANONICAL_POINTS exports exactly 12 points', () => {
|
|
assert.strictEqual(
|
|
ALL_12_POINTS.length,
|
|
12,
|
|
`CANONICAL_POINTS must have 12 entries (ADR-857 §"Loop Extension Points (the 12)"), got ${ALL_12_POINTS.length}`,
|
|
);
|
|
});
|
|
|
|
test('each of the 12 known point names is present in CANONICAL_POINTS', () => {
|
|
const expected = [
|
|
'discuss:pre', 'discuss:post',
|
|
'plan:pre', 'plan:post',
|
|
'execute:pre', 'execute:wave:pre', 'execute:wave:post', 'execute:post',
|
|
'verify:pre', 'verify:post',
|
|
'ship:pre', 'ship:post',
|
|
];
|
|
for (const p of expected) {
|
|
assert.ok(
|
|
ALL_12_POINTS.includes(p),
|
|
`Expected canonical point "${p}" to be in CANONICAL_POINTS`,
|
|
);
|
|
}
|
|
});
|
|
});
|
|
|
|
// ─────────────────────────────────────────────────────────────────────────────
|
|
// B2. [happy] CLI render-hooks E2E: exit 0, activeHooks:[], placeholder rendered
|
|
// ─────────────────────────────────────────────────────────────────────────────
|
|
|
|
describe('B2 — CLI loop render-hooks: exit 0, activeHooks:[], placeholder for all-caps-off project', () => {
|
|
let tmpDir;
|
|
|
|
before(() => {
|
|
tmpDir = makeProject(ALL_FALSE_CONFIG);
|
|
});
|
|
|
|
after(() => {
|
|
removeTmp(tmpDir);
|
|
tmpDir = null;
|
|
});
|
|
|
|
for (const point of ['plan:pre', 'execute:wave:post', 'ship:post']) {
|
|
test(`loop render-hooks ${point} → exit 0, activeHooks:[], non-empty rendered placeholder`, () => {
|
|
const { exitCode, output, error } = runCli(['loop', 'render-hooks', point], tmpDir);
|
|
|
|
assert.strictEqual(
|
|
exitCode,
|
|
0,
|
|
`Expected exit 0 for "loop render-hooks ${point}" with all-caps-off config; got ${exitCode}. stderr: ${error ?? ''}`,
|
|
);
|
|
|
|
// Must parse as JSON
|
|
let parsed;
|
|
try {
|
|
parsed = JSON.parse(output);
|
|
} catch (e) {
|
|
assert.fail(`CLI output for ${point} is not valid JSON: ${output.slice(0, 200)}`);
|
|
}
|
|
|
|
// activeHooks must be present and empty
|
|
assert.ok(
|
|
Array.isArray(parsed.activeHooks),
|
|
`${point}: activeHooks must be an array`,
|
|
);
|
|
assert.strictEqual(
|
|
parsed.activeHooks.length,
|
|
0,
|
|
`${point}: expected activeHooks:[] with all-caps-off config, got ${JSON.stringify(parsed.activeHooks)}`,
|
|
);
|
|
|
|
// rendered field must be a non-empty placeholder string (loop still renders output)
|
|
assert.ok(
|
|
typeof parsed.rendered === 'string' && parsed.rendered.length > 0,
|
|
`${point}: rendered must be a non-empty string, got ${JSON.stringify(parsed.rendered)}`,
|
|
);
|
|
|
|
// Genuine assertion: the placeholder contains the point name so it doesn't silently
|
|
// return a generic empty string detached from the requested point.
|
|
assert.ok(
|
|
parsed.rendered.includes(point),
|
|
`${point}: rendered placeholder must reference the point name "${point}", got: "${parsed.rendered}"`,
|
|
);
|
|
});
|
|
}
|
|
});
|
|
|
|
// ─────────────────────────────────────────────────────────────────────────────
|
|
// B3. [happy] Init bundles resolve with exit 0 and valid JSON with caps off
|
|
// ─────────────────────────────────────────────────────────────────────────────
|
|
|
|
describe('B3 — init bundles for 5-step loop entry seam: exit 0 and valid JSON with capabilities off', () => {
|
|
// Cases: the 5-step loop's main init entry points (those available without git)
|
|
const INIT_CASES = [
|
|
{
|
|
label: 'init plan-phase',
|
|
args: ['init', 'plan-phase', '--phase', '01-stub'],
|
|
// Required fields that prove the bundle is a real JSON object used by the loop
|
|
requiredFields: ['tdd_mode', 'phase_found', 'planning_exists'],
|
|
},
|
|
{
|
|
label: 'init execute-phase',
|
|
args: ['init', 'execute-phase', '--phase', '01-stub'],
|
|
requiredFields: ['tdd_mode', 'phase_found', 'config_exists'],
|
|
},
|
|
{
|
|
label: 'init verify-work',
|
|
args: ['init', 'verify-work', '--phase', '01-stub'],
|
|
requiredFields: ['phase_found', 'commit_docs'],
|
|
},
|
|
];
|
|
|
|
for (const { label, args, requiredFields } of INIT_CASES) {
|
|
describe(label, () => {
|
|
let tmpDir;
|
|
|
|
before(() => {
|
|
// Bare project with .planning/ but all caps off in config
|
|
tmpDir = makeProject(ALL_FALSE_CONFIG);
|
|
});
|
|
|
|
after(() => {
|
|
removeTmp(tmpDir);
|
|
tmpDir = null;
|
|
});
|
|
|
|
test(`${label} exits 0 with capabilities off`, () => {
|
|
const { exitCode, error } = runCli(args, tmpDir);
|
|
assert.strictEqual(
|
|
exitCode,
|
|
0,
|
|
`${label}: expected exit 0 with all-caps-off project, got ${exitCode}. stderr: ${error ?? ''}`,
|
|
);
|
|
});
|
|
|
|
test(`${label} returns parseable JSON with expected fields`, () => {
|
|
const { output } = runCli(args, tmpDir);
|
|
let parsed;
|
|
try {
|
|
parsed = JSON.parse(output);
|
|
} catch (e) {
|
|
assert.fail(`${label}: output is not valid JSON: ${output.slice(0, 200)}`);
|
|
}
|
|
assert.ok(
|
|
parsed && typeof parsed === 'object' && !Array.isArray(parsed),
|
|
`${label}: JSON must be a plain object`,
|
|
);
|
|
for (const field of requiredFields) {
|
|
assert.ok(
|
|
Object.prototype.hasOwnProperty.call(parsed, field),
|
|
`${label}: bundle must contain field "${field}", got keys: ${Object.keys(parsed).join(', ')}`,
|
|
);
|
|
}
|
|
});
|
|
|
|
test(`${label} returns parseable JSON with bare project (no config at all)`, () => {
|
|
const bareDir = makeProject(null); // no config.json
|
|
try {
|
|
const { exitCode, output, error } = runCli(args, bareDir);
|
|
assert.strictEqual(
|
|
exitCode,
|
|
0,
|
|
`${label}: expected exit 0 with bare project (no config), got ${exitCode}. stderr: ${error ?? ''}`,
|
|
);
|
|
let parsed;
|
|
try {
|
|
parsed = JSON.parse(output);
|
|
} catch (e) {
|
|
assert.fail(`${label}: bare project output is not valid JSON: ${output.slice(0, 200)}`);
|
|
}
|
|
assert.ok(
|
|
parsed && typeof parsed === 'object' && !Array.isArray(parsed),
|
|
`${label}: bare project JSON must be a plain object`,
|
|
);
|
|
} finally {
|
|
removeTmp(bareDir);
|
|
}
|
|
});
|
|
});
|
|
}
|
|
});
|
|
|
|
// ─────────────────────────────────────────────────────────────────────────────
|
|
// B4. [negative] Empty registry → activeHooks:[] at all 12 points
|
|
// ─────────────────────────────────────────────────────────────────────────────
|
|
|
|
describe('B4 — empty registry (byLoopPoint:{}) at all 12 points → activeHooks:[]', () => {
|
|
const EMPTY_REGISTRY = {
|
|
byLoopPoint: {},
|
|
capabilities: {},
|
|
configKeys: {},
|
|
configSchema: {},
|
|
commandFamilies: {},
|
|
};
|
|
|
|
test('loop tolerates a capability-less install: all 12 points return activeHooks:[]', () => {
|
|
const failures = [];
|
|
for (const point of ALL_12_POINTS) {
|
|
const result = resolveLoopHooks({
|
|
point,
|
|
registry: EMPTY_REGISTRY,
|
|
config: {},
|
|
});
|
|
assert.ok(
|
|
result && typeof result === 'object',
|
|
`${point}: result must be an object`,
|
|
);
|
|
assert.ok(
|
|
Array.isArray(result.activeHooks),
|
|
`${point}: activeHooks must be an array`,
|
|
);
|
|
if (result.activeHooks.length !== 0) {
|
|
failures.push({
|
|
point,
|
|
count: result.activeHooks.length,
|
|
capIds: result.activeHooks.map(h => h.capId),
|
|
});
|
|
}
|
|
}
|
|
assert.deepStrictEqual(
|
|
failures,
|
|
[],
|
|
`Empty registry: expected zero active hooks at all 12 points, got non-empty at: ${JSON.stringify(failures)}`,
|
|
);
|
|
});
|
|
|
|
test('empty registry does not throw for any of the 12 canonical points', () => {
|
|
for (const point of ALL_12_POINTS) {
|
|
assert.doesNotThrow(
|
|
() => resolveLoopHooks({ point, registry: EMPTY_REGISTRY, config: {} }),
|
|
`resolveLoopHooks must not throw for empty registry at point "${point}"`,
|
|
);
|
|
}
|
|
});
|
|
|
|
test('renderLoopHooks with empty activeHooks returns a non-empty placeholder string', () => {
|
|
const placeholder = renderLoopHooks({ point: 'plan:pre', activeHooks: [] });
|
|
assert.ok(
|
|
typeof placeholder === 'string' && placeholder.length > 0,
|
|
`renderLoopHooks must return a non-empty string for empty activeHooks, got: ${JSON.stringify(placeholder)}`,
|
|
);
|
|
// Specific value check — genuineness: this must change if the placeholder format changes
|
|
assert.strictEqual(
|
|
placeholder,
|
|
'_No active hooks at plan:pre._',
|
|
`renderLoopHooks placeholder must be "_No active hooks at plan:pre._", got: "${placeholder}"`,
|
|
);
|
|
});
|
|
});
|
|
|
|
// ─────────────────────────────────────────────────────────────────────────────
|
|
// B5. [BVA] One capability ON (tdd_mode) → its 2 points non-empty, 10 others empty
|
|
// ─────────────────────────────────────────────────────────────────────────────
|
|
|
|
describe('B5 — BVA: tdd_mode ON, all other caps OFF → additive: only tdd points active', () => {
|
|
// tdd contributes at plan:pre and execute:post (verified from capability registry)
|
|
const TDD_ACTIVE_POINTS = ['plan:pre', 'execute:post'];
|
|
const TDD_INACTIVE_POINTS = ALL_12_POINTS.filter(p => !TDD_ACTIVE_POINTS.includes(p));
|
|
const TDD_ON_CONFIG = buildTddOnlyConfig();
|
|
|
|
test('plan:pre has exactly 1 active hook and it belongs to tdd', () => {
|
|
const result = resolveLoopHooks({
|
|
point: 'plan:pre',
|
|
registry: realRegistry,
|
|
config: TDD_ON_CONFIG,
|
|
});
|
|
assert.ok(Array.isArray(result.activeHooks), 'activeHooks must be an array');
|
|
assert.strictEqual(
|
|
result.activeHooks.length,
|
|
1,
|
|
`plan:pre: expected 1 active hook (tdd), got ${result.activeHooks.length}: ${JSON.stringify(result.activeHooks.map(h => h.capId))}`,
|
|
);
|
|
assert.strictEqual(
|
|
result.activeHooks[0].capId,
|
|
'tdd',
|
|
`plan:pre: expected activeHooks[0].capId to be "tdd", got "${result.activeHooks[0].capId}"`,
|
|
);
|
|
});
|
|
|
|
test('execute:post has exactly 1 active hook and it belongs to tdd', () => {
|
|
const result = resolveLoopHooks({
|
|
point: 'execute:post',
|
|
registry: realRegistry,
|
|
config: TDD_ON_CONFIG,
|
|
});
|
|
assert.ok(Array.isArray(result.activeHooks), 'activeHooks must be an array');
|
|
assert.strictEqual(
|
|
result.activeHooks.length,
|
|
1,
|
|
`execute:post: expected 1 active hook (tdd gate), got ${result.activeHooks.length}: ${JSON.stringify(result.activeHooks.map(h => h.capId))}`,
|
|
);
|
|
assert.strictEqual(
|
|
result.activeHooks[0].capId,
|
|
'tdd',
|
|
`execute:post: expected activeHooks[0].capId to be "tdd", got "${result.activeHooks[0].capId}"`,
|
|
);
|
|
});
|
|
|
|
test('all 10 non-tdd points return activeHooks:[] even with tdd_mode ON', () => {
|
|
const failures = [];
|
|
for (const point of TDD_INACTIVE_POINTS) {
|
|
const result = resolveLoopHooks({
|
|
point,
|
|
registry: realRegistry,
|
|
config: TDD_ON_CONFIG,
|
|
});
|
|
assert.ok(Array.isArray(result.activeHooks), `${point}: activeHooks must be an array`);
|
|
if (result.activeHooks.length !== 0) {
|
|
failures.push({
|
|
point,
|
|
count: result.activeHooks.length,
|
|
capIds: result.activeHooks.map(h => h.capId),
|
|
});
|
|
}
|
|
}
|
|
assert.deepStrictEqual(
|
|
failures,
|
|
[],
|
|
`Expected 10 non-tdd points to be empty with tdd_mode ON, got non-zero at: ${JSON.stringify(failures)}`,
|
|
);
|
|
});
|
|
|
|
test('tdd hook at plan:pre is a contribution kind (not a gate or step)', () => {
|
|
const result = resolveLoopHooks({
|
|
point: 'plan:pre',
|
|
registry: realRegistry,
|
|
config: TDD_ON_CONFIG,
|
|
});
|
|
assert.strictEqual(
|
|
result.activeHooks.length,
|
|
1,
|
|
'Expected exactly 1 active hook at plan:pre with tdd ON',
|
|
);
|
|
assert.strictEqual(
|
|
result.activeHooks[0].kind,
|
|
'contribution',
|
|
`plan:pre tdd hook must be kind "contribution", got "${result.activeHooks[0].kind}"`,
|
|
);
|
|
});
|
|
|
|
test('tdd hook at execute:post is a gate kind (not a contribution or step)', () => {
|
|
const result = resolveLoopHooks({
|
|
point: 'execute:post',
|
|
registry: realRegistry,
|
|
config: TDD_ON_CONFIG,
|
|
});
|
|
assert.strictEqual(
|
|
result.activeHooks.length,
|
|
1,
|
|
'Expected exactly 1 active hook at execute:post with tdd ON',
|
|
);
|
|
assert.strictEqual(
|
|
result.activeHooks[0].kind,
|
|
'gate',
|
|
`execute:post tdd hook must be kind "gate", got "${result.activeHooks[0].kind}"`,
|
|
);
|
|
});
|
|
|
|
test('turning tdd_mode OFF restores both tdd points to activeHooks:[]', () => {
|
|
// Regression check: tdd_mode OFF → both previously-active points go back to empty
|
|
const tddOffConfig = buildAllFalseConfig(); // tdd_mode: false
|
|
for (const point of TDD_ACTIVE_POINTS) {
|
|
const result = resolveLoopHooks({
|
|
point,
|
|
registry: realRegistry,
|
|
config: tddOffConfig,
|
|
});
|
|
assert.strictEqual(
|
|
result.activeHooks.length,
|
|
0,
|
|
`${point}: expected activeHooks:[] with tdd_mode OFF, got ${JSON.stringify(result.activeHooks.map(h => h.capId))}`,
|
|
);
|
|
}
|
|
});
|
|
});
|