Files
msd-core/tests/ui-spec-inventory-provenance.test.cjs
Tom Boucher 15af0f5536 enhance(#3951): B6+B7 — widen two unreachable lint rules and make the guard ledger true (#3965)
* fix(#3951): two lint rules that could not reach the code they govern

B6 names two widenings. Measuring them first turned up a defect the criterion did
not know about, and refuted the reason it gave for one of them.

1. no-adhoc-markdown-parsing self-gates on its own filename.

   Lines 107-110 short-circuit create() to {} unless the path matches
   /(?:^|\/)src\/[^/]+\.cts$/. B6 says to widen the files: glob in
   eslint.config.mjs - but doing only that ships an INERT rule, because the gate
   still returns {} for every new path. Both halves have to change, and the gate
   is the load-bearing one.

   That same regex hides a live hole: [^/]+ is FLAT-ONLY, so it requires the file
   to sit directly in src/. The registered glob is src/**/*.cts, which includes
   subdirectories. 28 .cts files - health-diagnostic-rules/ (10),
   installer-migrations/ (11), observability/ (3), host-integration-adapters/ (2),
   vendor/ (2) - are inside the registered glob and silently skipped.

   Measured with the gate neutralized: 0 violations there today. The hole is
   hiding nothing right now, and is fixed anyway, because "no violations today" is
   not a property that keeps holding.

   The fix is not invented: require-subprocess-timeout.cjs:196 already carries the
   correct form of this guard, /(?:^|\/)src\/.*\.cts$/ with .*, one directory over.
   Checked the other 21 rules for the same bug - no-adhoc-regex-escape and
   no-private-binary-resolution short-circuit only to exempt their own seam file,
   which is the right shape, and no-crlf-fragile-split has no filename gate at
   all. This bug is unique to the one rule.

2. no-adhoc-regex-escape could not see the shape that actually occurs.

   Line 396 gated the whole UNSAFE-NEW-REGEXP arm on arg.type === 'Identifier'.
   Every check below it - the _SOURCE provenance check, the
   isSoleReturnOfOwnParameter shape - lives inside that branch, so
   new RegExp(obj['key']) and new RegExp(cfg.pattern) were never examined at all.
   Runtime data arrives as a property access far more often than as a bare
   identifier, which is exactly why this rule never fired on the #3477 ReDoS.

   Widened to MemberExpression, measured by AST walk across all five registered
   blocks rather than by grep. 27 sites, zero TSAsExpression:

     18  safe new RegExp(X.source, flags)  -> exempted, keyed strictly on the
         PROPERTY being `source`, never on the object. Keying on the object would
         wave through X.anything and buy nothing. B6 estimated ~10; that was an
         undercount.
      3  _SOURCE-suffixed constants reached through a required module namespace
         (phaseId.BRACKET_PHASE_TOKEN_SOURCE) -> the same provenance-exempt class
         the rule already recognizes for bare identifiers, extended to reach them.
         Without this the widening produces 3 false flags.
      6  real findings -> marked, each a test extracting a pattern from a shipped
         file at test time, where the runtime contract IS the product.

   Deliberately the NARROW MemberExpression form. The rule's own
   isSoleReturnOfOwnParameter doc comment records that an earlier broad
   "any non-literal identifier" heuristic produced ~25 false positives and was
   rejected; a re-run of the census after this change flags exactly the 6 above
   and nothing else.

Verified by execution, not by reading: the gate now accepts src/<subdir>/x.cts,
still accepts flat src/x.cts, and still exempts paths outside src/ - each pinned
by a test proven to fail against the old regex. build:lib, lint and lint:ci all
exit 0.

Refs #3951

Co-Authored-By: Claude Opus 5 <noreply@anthropic.com>

* fix(#3951): give no-adhoc-markdown-parsing its reach, and fix the 80 parses it finds

The rule self-gates on filename AND is registered on one glob, so widening either
half alone is inert. Both move here: the gate now accepts tests/**/*.cjs and
scripts/**/*.cjs alongside src/**/*.cts, and eslint.config.mjs registers it on the
same two.

A test pins that the gate and the registration AGREE, in both directions. The
original defect was a gate narrower than its registration; the failure mode of
this fix is a gate wider than its registration. Both are silent, so the test
asserts the pair rather than either half.

80 violations across 43 files, all in tests/, zero in scripts/. 70 are routed
through the existing seams - scanFencedBlocks, collectSection, stripFencedCode,
tokenizeHeadings from markdown-sectionizer; splitTableRow, parseMarkdownTable,
findTableWithColumns from markdown-table. Headerless STATE.md tables use
splitTableRow per line, because parseMarkdownTable needs a real delimiter row.

10 are suppressed, 12.5%, well under the third that would have meant the rule is
mis-scoped for tests/ rather than the tests carrying debt. Each names its reason:
three regression guards (#3873 / bug-#21) are deliberately independent of the
generator's own fence handling, and routing them through the seam would have them
test the generator against itself; one is a negative-text probe that extracts
nothing; six are a shell-pipe-to-jq detector whose regex coincidentally matches the
table fingerprint and is not markdown parsing at all.

All ten sit in tests whose subject is .md content, which is normally a reason to
prefer the seam. The marker used is allow-adhoc-markdown, distinct from
no-source-grep's allow-test-rule, and lint:ci's lint-allow-test-rule-refs reports
the same 280/280 unverified count as before - checked rather than assumed, because
those two markers are easy to conflate.

The widening earned its keep immediately: it found a test that passed for the
wrong reason.

  tests/config-field-docs.test.cjs asserted notEqual(<cell>, '600') against the
  TYPE column instead of the DEFAULT column. notEqual('number', '600') is true
  forever, so the guard against workflow.subagent_timeout regressing to the old
  seconds default could never fire. docs/CONFIGURATION.md:434 is
  `| workflow.subagent_timeout | number | 300000 | ... |`, so the default is cell
  index 2; the assertion is now row-scoped through splitTableRow and reads 300000.

That is the argument for the widening in one case: the violation was invisible to
lint, the suite was green, and the assertion was vacuous. A rule that cannot reach
a file cannot tell you the file is lying.

Not fixed here, and recorded rather than assumed: #3426/#3239 are NOT reachable by
this widening. tests/package-legitimacy-gate.test.cjs yields zero violations even
with the gate bypassed - its hand-rolled scans are real, but built from line
filters and split('|') rather than the regex-literal fingerprints this rule
detects. They need new detectors. The epic assumed a wider glob would catch them.

build:lib, lint and lint:ci all exit 0; the post-fix census across tests/** and
scripts/** is 0 violations.

Refs #3951

Co-Authored-By: Claude Opus 5 <noreply@anthropic.com>

* fix(#3951): B7 — and #3356's defects were still live in the code

B7 asks that each closed child be driven fail-first with a behavioral identity
test at the CONSUMER's output. Four of eleven children had no test citing their
issue number. Auditing them by BEHAVIOR rather than by number-grep changed the
answer for three of the four.

#3364 and #2540 — traceability only. Both were implemented by #3941 and their
consumer-output tests exist and were shown failing-first; neither cited its
originating issue, so an audit that greps for the number reports them uncovered.
Tagged the specific asserting test in each file, following the citation form those
files already use.

#3372 — covered, but only at helper level, and the triage narrowed it. Of the four
commands the issue names, only estimate-cli's collectCalibrationSamples actually
enumerates phase dirs from disk; smart-entry, audit and roadmap-upgrade derive from
ROADMAP/body text and never reach the sentinel path, so they are benign by
construction and were left alone rather than "fixed" into churn. The existing #3882
rows asserted the helper's return value. Added a consumer-output test driving
`query estimate-calibrate` and asserting sample_count and the persisted document.
RED proof: reverted collectCalibrationSamples to a raw readdirSync and ran the real
CLI - sample_count 3, sentinel leaked; restored - sample_count 2.

#3356 — NOT covered, and BOTH halves of the defect were still live in source. The
issue is closed; the bug was not fixed. Fixed here rather than writing tests that
document a bug as correct.

  Defect 1, the contradicted row. quick.md:627 claimed
  `quick-tasks-append` performs "the equivalent write" to the Step 7c row. It did
  not: the `#` cell was a positional ordinal and `Directory` read `—`, because the
  route had no way to receive a quick id or task directory. Added OPTIONAL
  `--quick-id` / `--slug` / `--directory`. A caller with neither - fast.md, the
  original #2133 caller - omits them and gets the byte-identical prior row, so
  nothing existing changes. A caller that HAS a real id and directory now gets the
  canonical row quick.md:632 renders. The false-equivalence sentence itself is
  corrected rather than left to mislead the next reader.

  Defect 2, the forced re-derive. The route called readModifyWriteStateMd with no
  options, so a body-only append to the Quick Tasks table triggered a full
  re-derive of the disk-derived progress.* frontmatter. Every other body-only
  writer passes { resync: false } - src/state.cts's own docstring prescribes it -
  and this route was the lone outlier. RED proof: reverted the option, seeded a
  project with 2 real phase dirs and a curated total_phases of 25, ran
  quick-tasks-append; total_phases collapsed to 2. Restored; it stayed 25.

That second one is the shape this epic exists to close: a silent write that
replaces curated state with a re-derivation nobody asked for, exit 0 throughout.

build:lib, lint and lint:ci all exit 0.

Refs #3951

Co-Authored-By: Claude Opus 5 <noreply@anthropic.com>

* docs(#3951): amend B6's ledger to what was measured, and document the new flags

The ADR gains a ledger amendment in its own correction style - the sixth wrong
premise it records, found the same way as the other five, by measuring before
building.

B6 says the net guard count must fall. It rose: 62 -> 69, +7, measured from the
epic's filing commit to origin/next. The attribution is the point, though. Five of
the seven came from PRs unrelated to this epic, one was added by a phase of it, and
the epic did retire something sub-file - #3884 removed a detector with an explicit
"net: -1 detector, 0 added" ledger. Every named casualty is load-bearing, two
already carry retractions in this same document, and a sweep of all 22 rules plus
every scripts/lint-* found no provably dead guard. There is no honest way to make
the count fall; forcing it would trade coverage for a number, which is the Goodhart
outcome Decision 6 exists to prevent.

The amendment also records that B6's own prescribed fix for one widening was inert.
no-adhoc-markdown-parsing self-gates on its filename, so widening only the files:
glob - which is what the criterion says to do - ships a rule that still returns {}
for every new path. And #3426/#3239 are not reachable by that widening at all;
their scans use line filters and split('|'), not the regex fingerprints the rule
detects. The roster row tracked them against the wrong mechanism.

Three roster rows updated from aspiration to fact: the two widenings are DONE with
their measured counts, and lint-phase-enumeration-drift is marked RETAINED rather
than "expected casualty - verify before retiring", because Phase 5 verified it and
kept it.

The rule Decision 6 should carry forward is stated plainly: a guard ledger is a
claim about COVERAGE, not about COUNT. "Net count must fall" is measurable and
wrong. "Every guard is reachable, and each retirement names what makes its defect
unrepresentable" is the property that was actually wanted.

CLI-TOOLS.md documents the optional --quick-id/--slug/--directory flags and says
plainly that omitting them keeps the pre-#3356 row byte-identical, plus that the
append no longer re-derives progress frontmatter.

New features fragment (id 3951); FEATURES.md regenerated rather than hand-edited.
Changeset is Changed, pr:0 pending backfill.

Refs #3951

Co-Authored-By: Claude Opus 5 <noreply@anthropic.com>

* test(#3951): correct four rows that pinned the lint rule's old narrow reach

The remote suite came back RED with 5 failures, all in tests/eslint-rules.test.cjs.
They are stale tests, not a regression: four rows assert that
no-adhoc-markdown-parsing is inert outside src/*.cts, which is exactly the
contract this deliverable changes.

Confirmed by reading rather than inferred from the names - the row at :1981 used
filename: 'tests/some.test.cjs' and filename: 'scripts/helper.cjs', the two roots
the rule now covers on purpose.

Worth recording WHY local gates missed this. npm run lint and lint:ci were green,
and the touched test files passed standalone. Lint only reports violations in real
files; these rows assert the rule's REACH using synthetic RuleTester filenames, so
nothing but the full suite could see them. Local green on a rule change says
nothing about the rule's own tests.

Each row is rewritten with BOTH halves rather than flipped from valid to invalid:

  - the same fingerprint under tests/ or scripts/ is now flagged, with the right
    messageId
  - the negative space is preserved - the same fingerprint under a path outside
    all three roots (gsd-core/bin/lib/foo.cjs) is still NOT flagged

The second half is the one that matters. Without it the rule has no boundary and
nothing would catch an over-wide gate later, which is the mirror image of the bug
this deliverable just fixed.

Each row is renamed to state the current contract; the old names said
"non-src/*.cts ... is not flagged" and would have been actively misleading once
the bodies changed.

Proven to test the widening rather than restate it: every flagged half was run
against HEAD~2's pre-widening rule and does NOT fire there, then against the
current rule and does. 12/12 on that probe; the full file is 178/178.

Swept for the same staleness elsewhere and found none.
require-subprocess-timeout's own "inert outside src/*.cts" row is untouched -
that rule's gate was not widened here - and no-adhoc-regex-escape's test file
already carries correctly-targeted rows.

Refs #3951

Co-Authored-By: Claude Opus 5 <noreply@anthropic.com>

* test(#3951): acknowledge the quick.md growth the attribution guard reported

The full suite came back RED with one failure, and it is mine:

  1 file(s) grew without an acknowledgment:
    quick.md grew 364 bytes

gsd-core/workflows/quick.md is runtime-loaded emitted content, so correcting
its false 'performs the equivalent write' claim trips emitted-attribution by
construction. This is the acknowledgment, not a workaround - there is nothing
to regenerate.

The fragment names ONE path, which is the only one the guard reported. The four
spent acknowledgments it also listed (audit-uat, plan-phase, progress, review)
belong to other fragments whose ripple the base already absorbs; they are inert,
not failures, and are deliberately NOT copied here - naming paths I did not
change would make this record false in the other direction.

Byte figure corrected before committing: the guard reported 37220 -> 37584
(+364), but origin/next has since moved and quick.md is 37232 there now, so the
measured delta is +352. The reason text says so and names the base as a moving
figure rather than pinning a number that is already stale.

Refs #3951

Co-Authored-By: Claude Opus 5 <noreply@anthropic.com>

* test(#3951): move the quick.md growth ack to a trailer, delete the obsolete fragment

The acknowledgment mechanism changed under this branch. Merging next brought in
the redesign - it also deleted .github/workflows/ack-fragment-sweep.yml, which
was in the merge status and which I did not register at the time - and the guard
now says so directly:

  Add a trailer to a commit in this PR (never a new file).
    Emitted-Drift-Ack-Growth: quick.md - <why this growth is deliberate>

So tests/emitted-drift-acks/3951-quick-append-equivalence.json is obsolete on
arrival. A fragment file is no longer read by anything, and leaving it would be a
dead record that looks like an active one. It is deleted here rather than kept
"just in case".

The byte figure moved again with the merge: 37232 -> 37596, +364. The earlier
fragment said +352, measured before the merge auto-merged quick.md itself. The
trailer carries no number, which is the better design - the figure was stale
twice in two attempts.

Refs #3951

Emitted-Drift-Ack-Growth: quick.md — #3356/#3951 replaces a false claim with an accurate one. Line 627 said the `quick-tasks-append` shortcut "performs the equivalent write" to the Step 7c row rendered above it; it did not, and that was the documented half of #3356 — with no quick id or task directory the route emitted a positional ordinal in `#` and an em-dash in `Directory`, a visibly different row. The corrected sentence has to carry three facts the original elided: what the shortcut actually writes when it has neither input, that this is honest behavior for its real caller (`fast.md`, which has neither), and how a caller with both now gets the byte-identical canonical row via the new optional `--quick-id`/`--slug`/`--directory` flags. Prose is the product here — an executing agent reads this line to decide whether the shortcut is safe for its case, and a shorter correction would either drop the flags (leaving the reader unable to act on the fix) or drop the limitation (recreating the false claim in gentler words).
Co-Authored-By: Claude Opus 5 <noreply@anthropic.com>

* chore(#3951): backfill changeset pr number

Refs #3951

Co-Authored-By: Claude Opus 5 <noreply@anthropic.com>

---------

Co-authored-by: sim <sim@local>
Co-authored-by: Claude Opus 5 <noreply@anthropic.com>
2026-08-27 23:10:49 -04:00

742 lines
34 KiB
JavaScript
Raw Blame History

This file contains ambiguous Unicode characters
This file contains Unicode characters that might be confused with other characters. If you think that this is intentional, you can safely ignore this warning. Use the Escape button to reveal them.
'use strict';
/**
* UI-SPEC component-inventory provenance (#2845).
*
* Two contracts are under test, and both are SHARED FORMATS spread across surfaces
* with no generator — the `DEFECT.GENERATIVE-FIX-DIVERGENCE` class:
*
* 1. The gsd-ui-checker DIMENSION ROSTER, asserted independently on twelve surfaces
* (eight English, four translated). Adding Dimension 7 is only correct if every
* one of them moves together; a count-only guard would miss a relabel, and a
* guard that reads only the real tree never executes its own failure branch.
* Every parity assertion below is therefore paired with a synthetic MUTATION case.
*
* 2. The PROVENANCE-LINE GRAMMAR, emitted by `gsd-core/templates/UI-SPEC.md`
* (`## Component Inventory`) and consumed by `agents/gsd-ui-checker.md`
* (Dimension 7). One format, two surfaces, opposite directions — the two must
* quote it byte-identically or the checker keys on a shape the template never emits.
*
* The units under test are the shipped markdown documents themselves: their text IS
* what the runtime loads (CONTRIBUTING.md — `source-text-is-the-product`). Structure is
* asserted on parsed, typed records (`{ n, label }`, token sets, violation objects)
* rather than on raw substrings, per CONTRIBUTING.md's "Prohibited: Raw Text Matching".
*
* Three assertions are a deliberate exception: Dimension 7's allowlist-downgrade
* wording, its not-applicable PASS clause, and the template's non-exhaustive clause are
* CONTRACT PROSE — the sentence itself is the deliverable an agent reads at runtime, so
* there is no typed IR to assert on instead. That is the `source-text-is-the-product`
* category, not an escape from the rule.
*
* They carry NO `allow-test-rule` marker, deliberately and against first instinct.
* `no-source-grep` only inspects reads of `.cjs`/`.cts`/`.js`/`.mjs`/`.mts`/`.ts` paths,
* never `.md` ones, so a marker here suppresses nothing: it lands in the gate's
* "unverified" bucket, which is ceilinged. Measured on 2026-08-21 — adding three markers
* took that count 280 -> 281 and FAILED `scripts/lint-allow-test-rule-refs.cjs`. Raising
* the ceiling for markers that suppress nothing is the "just bump the baseline" weakening
* the ratchet exists to prevent, so the honest answer is no marker and this note.
*
* See https://github.com/open-gsd/gsd-core/issues/2845
*/
const { describe, test } = require('node:test');
const assert = require('node:assert/strict');
const fs = require('node:fs');
const path = require('node:path');
const fc = require('fast-check');
const { splitTableRow } = require('../gsd-core/bin/lib/markdown-table.cjs');
const ROOT = path.join(__dirname, '..');
// ─── Shipped surfaces ─────────────────────────────────────────────────────────
const SURFACE = Object.freeze({
CHECKER: 'agents/gsd-ui-checker.md',
RESEARCHER: 'agents/gsd-ui-researcher.md',
TEMPLATE: 'gsd-core/templates/UI-SPEC.md',
WORKFLOW: 'gsd-core/workflows/ui-phase.md',
FEATURES: 'docs/FEATURES.md',
HOWTO: 'docs/how-to/design-a-ui-phase.md',
FEATURES_JA: 'docs/ja-JP/FEATURES.md',
HOWTO_JA: 'docs/ja-JP/how-to/design-a-ui-phase.md',
FEATURES_ZH: 'docs/zh-CN/FEATURES.md',
HOWTO_ZH: 'docs/zh-CN/how-to/design-a-ui-phase.md',
FEATURES_KO: 'docs/ko-KR/FEATURES.md',
HOWTO_KO: 'docs/ko-KR/how-to/design-a-ui-phase.md',
HOWTO_PT: 'docs/pt-BR/how-to/design-a-ui-phase.md',
PROBE_REFERENCE: 'gsd-core/references/ui-consideration-probe.md',
});
function readShipped(rel) {
const abs = path.join(ROOT, rel);
return fs.existsSync(abs) ? fs.readFileSync(abs, 'utf8') : '';
}
/** CRLF-normalize before any line splitting — a raw `\n` split leaves `\r` on every
* captured label and makes a Windows checkout parse differently (CRLF bug class). */
const lf = (text) => String(text == null ? '' : text).replace(/\r\n/g, '\n');
// ─── Parsers → typed IR ───────────────────────────────────────────────────────
/** Line-oriented scan yielding `{ n, label }` for every line matching `re`.
* `skipFenced` drops lines inside ``` blocks — a `## Dimension 8:` shown as an
* EXAMPLE inside a fence is documentation, not a roster entry. The verdict block
* is itself fenced, so its parser must NOT skip fences. */
function matchLines(text, re, { skipFenced = false } = {}) {
const out = [];
let inFence = false;
for (const line of lf(text).split('\n')) {
if (skipFenced && /^\s*```/.test(line)) { inFence = !inFence; continue; }
if (inFence) continue;
const m = re.exec(line);
if (m) out.push({ n: Number(m[1]), label: m[2].trim() });
}
return out;
}
/** `## Dimension <N>: <Label>` — the canonical roster. */
function parseDimensionHeadings(text) {
return matchLines(text, /^##\s+Dimension\s+(\d+):\s*(\S.*?)\s*$/, { skipFenced: true });
}
/** `Dimension <N> — <Label>: {PASS / FLAG / BLOCK}` — the verdict block (itself fenced). */
function parseVerdictBlock(text) {
return matchLines(text, /^Dimension\s+(\d+)\s*[—-]\s*(\S.*?):\s*\{/);
}
/** `- [ ] Dimension <N> <Label>: PASS` — the template's Checker Sign-Off. */
function parseSignOff(text) {
return matchLines(text, /^-\s+\[\s*\]\s+Dimension\s+(\d+)\s+(\S.*?):\s*PASS\s*$/);
}
/** `| <N> <Label> | {PASS/FLAG} | … |` — the structured-return dimension tables.
* Two such tables ship (VERIFIED and ISSUES FOUND), so rows are de-duplicated. */
function parseReturnTableRows(text) {
const seen = new Map();
for (const line of lf(text).split('\n')) {
if (!line.trim().startsWith('|')) continue;
const cells = splitTableRow(line);
const m = cells.length > 0 ? /^(\d+)\s+(\S.*?)\s*$/.exec(cells[0]) : null;
if (m) seen.set(`${m[1]}|${m[2]}`, { n: Number(m[1]), label: m[2] });
}
return [...seen.values()];
}
/** `1. **<Label>** — …` inside a "Validation Dimensions" list. Numerals stay ASCII in
* every locale, so this parses the translated lists too. */
function parseNumberedDimensionList(text) {
const out = [];
for (const line of lf(text).split('\n')) {
const m = /^(\d+)\.\s+\*\*(\S.*?)\*\*/.exec(line);
if (m) out.push({ n: Number(m[1]), label: m[2] });
}
return out;
}
const WORD_NUMERAL = Object.freeze({
five: 5, six: 6, seven: 7, eight: 8, nine: 9, ten: 10,
seis: 6, sete: 7, oito: 8,
五: 5, 六: 6, 七: 7, 八: 8, 九: 9, 十: 10,
여섯: 6, 일곱: 7, 여덟: 8,
});
/** A numeral introduced by "other" or "remaining" is a BACK-REFERENCE to the rest of a set
* — "the same as the other six dimensions" — never a claim about the set's SIZE. Counting
* one turned `next` red on dacae9273 against entirely correct documentation, and a drift
* guard that fires on accurate prose is a false positive, which is how guards end up
* disabled.
*
* Declared once and shared by both English scans so the two can never drift apart — the
* same divergence class this whole suite exists to catch. `\b` anchors it to the whole
* word, so "another six dimensions" (a real claim about a second set) still counts, and
* both scans apply it case-insensitively: a sentence-initial "Other six dimensions…" is
* the same back-reference as a mid-sentence one.
*
* Known limits: the exclusion is English-only — the ja/zh/ko/pt patterns have no
* equivalent, because translated docs here are CORRECTED to match English rather than
* authored, so there is no instance to model the grammar on. And it cannot tell a
* back-reference from a genuine total that happens to open with the same word ("Other 6
* dimensions were added"); prose alone does not disambiguate those, and the false-negative
* is the safer side of that trade for a guard whose failure mode is being switched off. */
const NOT_A_BACK_REFERENCE = String.raw`(?<!\b(?:other|remaining)\s)`;
/** Every numeric "N dimensions" claim in `text`, in any of the shipped languages.
* Returns plain numbers — the typed IR the parity function consumes.
*
* `excludeBackReferences: false` reproduces the pre-fix behavior. It exists so a test can
* prove, against the real shipped files, that the exclusion is load-bearing rather than
* decorative — see the dacae9273 regression block. */
function parseDeclaredCounts(text, { excludeBackReferences = true } = {}) {
const body = lf(text);
const found = [];
const push = (v) => { if (Number.isInteger(v)) found.push(v); };
const scan = (re, take) => { for (const m of body.matchAll(re)) take(m); };
const guard = excludeBackReferences ? NOT_A_BACK_REFERENCE : '';
// "6 dimensions", "6 design quality dimensions", "6 Validation Dimensions"
scan(new RegExp(`${guard}(\\d+)\\s+(?:[A-Za-z][A-Za-z-]*\\s+){0,3}?[Dd]imensions?\\b`, 'gi'),
(m) => push(Number(m[1])));
// "six dimensions", "six quality dimensions"
scan(new RegExp(`${guard}\\b(five|six|seven|eight|nine|ten)\\s+(?:[A-Za-z][A-Za-z-]*\\s+){0,3}?dimensions\\b`, 'gi'),
(m) => push(WORD_NUMERAL[m[1].toLowerCase()]));
// "Dimensions: 6/6 passed"
scan(/Dimensions:\s*(\d+)\/(\d+)/g, (m) => { push(Number(m[1])); push(Number(m[2])); });
// ja: "6つの次元", "6つのバリデーション次元", "6 つの側面"
scan(/(\d+)\s*つの[^\s。、]*?(?:次元|側面)/g, (m) => push(Number(m[1])));
// zh: "6 个维度"
scan(/(\d+)\s*个[^\s::,,。]{0,6}?维度/g, (m) => push(Number(m[1])));
// zh: "六个维度"
scan(/([五六七八九十])\s*个[^\s::,,。]{0,6}?维度/g, (m) => push(WORD_NUMERAL[m[1]]));
// ko: "7개 차원", "7가지 유효성 검사 차원"
scan(/(\d+)\s*(?:개|가지)\s*[^\n]{0,12}?차원/g, (m) => push(Number(m[1])));
// ko: "일곱 가지 차원"
scan(/(여섯|일곱|여덟)\s*가지\s*[^\n]{0,12}?차원/g, (m) => push(WORD_NUMERAL[m[1]]));
// pt: "7 dimensões"
scan(/(\d+)\s+dimens(?:ão|ões)/g, (m) => push(Number(m[1])));
// pt: "sete dimensões"
scan(/\b(seis|sete|oito)\s+dimens(?:ão|ões)\b/gi, (m) => push(WORD_NUMERAL[m[1].toLowerCase()]));
return found;
}
/** The `##`/`###`/`####` section whose body contains `needle`. */
function sectionContaining(text, needle) {
const lines = lf(text).split('\n');
const hit = lines.findIndex((l) => l.includes(needle));
if (hit === -1) return '';
const isHeading = (l) => /^#{2,4}\s/.test(l);
let start = hit;
while (start > 0 && !isHeading(lines[start])) start -= 1;
let end = hit + 1;
while (end < lines.length && !isHeading(lines[end])) end += 1;
return lines.slice(start, end).join('\n');
}
const PROVENANCE_TOKEN = Object.freeze({
COMMAND: 'command',
COUNT: 'count',
VERSION: 'version',
DATE: 'date',
});
const REQUIRED_PROVENANCE_TOKENS = Object.freeze([
PROVENANCE_TOKEN.COMMAND, PROVENANCE_TOKEN.COUNT,
PROVENANCE_TOKEN.VERSION, PROVENANCE_TOKEN.DATE,
]);
/** The canonical provenance grammar line and the placeholder tokens it carries.
* `null` when the section states no grammar at all. */
function parseProvenanceGrammar(text) {
const line = lf(text).split('\n').map((l) => l.trim())
.find((l) => /^Enumerated by\b/.test(l));
if (!line) return null;
const tokens = new Set();
if (/`<command>`/.test(line)) tokens.add(PROVENANCE_TOKEN.COMMAND);
if (/<N>\s+components/.test(line)) tokens.add(PROVENANCE_TOKEN.COUNT);
if (/<package>@<version>/.test(line)) tokens.add(PROVENANCE_TOKEN.VERSION);
if (/<YYYY-MM-DD>/.test(line)) tokens.add(PROVENANCE_TOKEN.DATE);
return { line, tokens };
}
/** The "could not enumerate" shape that must live in the SAME slot (#2845 A3). */
function parseCannotEnumerateShape(text) {
return lf(text).split('\n').map((l) => l.trim())
.find((l) => /^Could not enumerate:\s*<reason>/.test(l)) || null;
}
/** A fenced ```yaml example-issue block, as flat `key: value` records. */
function parseYamlExampleBlocks(text) {
const blocks = [];
let current = null;
for (const line of lf(text).split('\n')) {
if (/^```ya?ml\s*$/.test(line)) { current = {}; continue; }
if (current && /^```\s*$/.test(line)) { blocks.push(current); current = null; continue; }
if (current) {
const m = /^([a-z_]+):\s*(.*)$/.exec(line);
if (m) current[m[1]] = m[2].replace(/^"(.*)"$/, '$1').trim();
}
}
return blocks;
}
/** Which verdict tiers a dimension section declares criteria for. */
function parseVerdictTiers(section) {
const tiers = new Set();
for (const line of lf(section).split('\n')) {
const m = /^\*\*(BLOCK|FLAG|PASS) if:\*\*/.exec(line.trim());
if (m) tiers.add(m[1]);
}
return tiers;
}
// ─── The parity function (typed IR in, violations out) ────────────────────────
const rosterKey = (d) => `${d.n}|${d.label}`;
const rosterFingerprint = (roster) => roster.map(rosterKey).join(' ');
/**
* @returns {{kind:string, surface:string}[]} — empty when every surface agrees with
* `canonical`. Never throws; a surface that declares no count at all is itself a
* violation, because a pattern that went stale is how a parity guard turns vacuous.
*/
function checkRosterParity({ canonical, rosterSurfaces = [], countSurfaces = [] }) {
const violations = [];
if (canonical.length === 0) {
violations.push({ kind: 'empty-roster', surface: 'canonical' });
return violations;
}
const seen = new Set();
canonical.forEach((d, i) => {
if (d.n !== i + 1) violations.push({ kind: 'non-contiguous', surface: 'canonical', at: d.n });
if (seen.has(d.n)) violations.push({ kind: 'duplicate', surface: 'canonical', at: d.n });
seen.add(d.n);
});
const want = rosterFingerprint(canonical);
for (const s of rosterSurfaces) {
if (rosterFingerprint(s.roster) !== want) {
violations.push({ kind: 'roster-mismatch', surface: s.name, found: s.roster, expected: canonical });
}
}
for (const s of countSurfaces) {
if (s.counts.length === 0) {
violations.push({ kind: 'no-count-found', surface: s.name });
continue;
}
for (const c of s.counts) {
if (c !== canonical.length) {
violations.push({ kind: 'count-mismatch', surface: s.name, found: c, expected: canonical.length });
}
}
}
return violations;
}
// ─── Real-tree reads ──────────────────────────────────────────────────────────
const checker = readShipped(SURFACE.CHECKER);
const researcher = readShipped(SURFACE.RESEARCHER);
const template = readShipped(SURFACE.TEMPLATE);
const workflow = readShipped(SURFACE.WORKFLOW);
const CANONICAL = parseDimensionHeadings(checker);
const EXPECTED_DIMENSIONS = 7;
const DIMENSION_7_LABEL = 'Inventory Provenance';
const featuresRegion = (rel) => sectionContaining(readShipped(rel), 'REQ-UI-03');
// ─── Suites ───────────────────────────────────────────────────────────────────
describe('#2845 — gsd-ui-checker dimension roster', () => {
test('the checker declares a contiguous 1..7 roster whose 7th is Inventory Provenance', () => {
assert.equal(CANONICAL.length, EXPECTED_DIMENSIONS);
assert.deepEqual(CANONICAL.map((d) => d.n), [1, 2, 3, 4, 5, 6, 7]);
assert.deepEqual(CANONICAL[6], { n: 7, label: DIMENSION_7_LABEL });
});
test('the verdict block lists exactly the roster, same numbers and same labels', () => {
assert.deepEqual(parseVerdictBlock(checker), CANONICAL);
});
test('every roster dimension appears in the structured-return tables', () => {
const rowKeys = new Set(parseReturnTableRows(checker).map(rosterKey));
for (const d of CANONICAL) {
assert.ok(rowKeys.has(rosterKey(d)), `return tables omit "${d.n} ${d.label}"`);
}
});
test('the template Checker Sign-Off lists exactly the roster', () => {
assert.deepEqual(parseSignOff(template), CANONICAL);
});
});
describe('#2845 — cross-surface dimension-count parity (12 surfaces)', () => {
// Every surface that states a UI-checker dimension count, in any shipped language.
// ko-KR, pt-BR and the probe reference were MISSED on the first pass of #2845 and
// shipped stale — a guard that omits a surface is exactly as blind as no guard.
const COUNT_SURFACES = [
SURFACE.CHECKER, SURFACE.RESEARCHER, SURFACE.WORKFLOW, SURFACE.PROBE_REFERENCE,
SURFACE.FEATURES, SURFACE.HOWTO,
SURFACE.FEATURES_JA, SURFACE.HOWTO_JA,
SURFACE.FEATURES_ZH, SURFACE.HOWTO_ZH,
SURFACE.FEATURES_KO, SURFACE.HOWTO_KO,
SURFACE.HOWTO_PT,
];
test('no shipped surface still declares a stale dimension count', () => {
const countSurfaces = COUNT_SURFACES.map((rel) => ({
name: rel,
// FEATURES.md is a whole-product doc; scope it to the UI Design Contract section
// so gsd-plan-checker's own dimension counts elsewhere are not swept in.
counts: parseDeclaredCounts(
rel.endsWith('FEATURES.md') ? featuresRegion(rel) : readShipped(rel),
),
}));
const violations = checkRosterParity({ canonical: CANONICAL, countSurfaces });
assert.deepEqual(violations, [], JSON.stringify(violations, null, 2));
});
test('each FEATURES.md locale numbers its validation-dimension list 1..7', () => {
for (const rel of [SURFACE.FEATURES, SURFACE.FEATURES_JA, SURFACE.FEATURES_ZH, SURFACE.FEATURES_KO]) {
const list = parseNumberedDimensionList(featuresRegion(rel));
assert.deepEqual(list.map((d) => d.n), [1, 2, 3, 4, 5, 6, 7], `${rel} dimension list`);
}
});
test("the English FEATURES.md list uses the checker's own labels", () => {
assert.deepEqual(parseNumberedDimensionList(featuresRegion(SURFACE.FEATURES)), CANONICAL);
});
});
describe('#2845 — parity guard non-vacuity (mutation cases)', () => {
const canonical = [
{ n: 1, label: 'Copywriting' }, { n: 2, label: 'Visuals' }, { n: 3, label: 'Color' },
{ n: 4, label: 'Typography' }, { n: 5, label: 'Spacing' }, { n: 6, label: 'Registry Safety' },
{ n: 7, label: DIMENSION_7_LABEL },
];
const kinds = (v) => v.map((x) => x.kind);
test('accepts the aligned roster at the limit (7 on every surface)', () => {
assert.deepEqual(checkRosterParity({
canonical,
rosterSurfaces: [{ name: 'verdict', roster: canonical }],
countSurfaces: [{ name: 'docs', counts: [7, 7] }],
}), []);
});
test('fails at limit-1 — a surface still declaring 6', () => {
const v = checkRosterParity({ canonical, countSurfaces: [{ name: 'docs', counts: [6] }] });
assert.deepEqual(kinds(v), ['count-mismatch']);
assert.equal(v[0].found, 6);
assert.equal(v[0].surface, 'docs');
});
test('fails at limit+1 — a surface declaring 8', () => {
const v = checkRosterParity({ canonical, countSurfaces: [{ name: 'docs', counts: [8] }] });
assert.deepEqual(kinds(v), ['count-mismatch']);
assert.equal(v[0].found, 8);
});
test('fails when a surface stops declaring any count at all (stale pattern)', () => {
const v = checkRosterParity({ canonical, countSurfaces: [{ name: 'docs', counts: [] }] });
assert.deepEqual(kinds(v), ['no-count-found']);
});
test('fails when a dimension is dropped from one roster surface', () => {
const v = checkRosterParity({
canonical,
rosterSurfaces: [{ name: 'verdict', roster: canonical.slice(0, 6) }],
});
assert.deepEqual(kinds(v), ['roster-mismatch']);
});
test('fails on a label that drifts on one surface only — a count-only guard would pass this', () => {
const drifted = canonical.map((d) => (d.n === 7 ? { n: 7, label: 'Inventory Sourcing' } : d));
const v = checkRosterParity({ canonical, rosterSurfaces: [{ name: 'verdict', roster: drifted }] });
assert.deepEqual(kinds(v), ['roster-mismatch']);
// and the count-only lane is genuinely blind to it, which is why both lanes exist
assert.deepEqual(
checkRosterParity({ canonical, countSurfaces: [{ name: 'docs', counts: [drifted.length] }] }),
[],
);
});
test('fails on a non-contiguous roster', () => {
const gapped = [...canonical.slice(0, 4), { n: 6, label: 'Registry Safety' }];
assert.ok(kinds(checkRosterParity({ canonical: gapped })).includes('non-contiguous'));
});
test('fails on a duplicated dimension number', () => {
const duped = [...canonical.slice(0, 6), { n: 6, label: DIMENSION_7_LABEL }];
assert.ok(kinds(checkRosterParity({ canonical: duped })).includes('duplicate'));
});
test('an empty roster is a violation, not a silent pass', () => {
assert.deepEqual(kinds(checkRosterParity({ canonical: [] })), ['empty-roster']);
});
});
describe('#2845 — parsers are total and newline-agnostic', () => {
const parsers = [
parseDimensionHeadings, parseVerdictBlock, parseSignOff,
parseReturnTableRows, parseNumberedDimensionList, parseDeclaredCounts,
];
test('CRLF text parses identically to LF text on every real surface', () => {
for (const text of [checker, template, workflow, researcher]) {
const crlf = text.replace(/\r\n/g, '\n').replace(/\n/g, '\r\n');
for (const parse of parsers) {
assert.deepEqual(parse(crlf), parse(text), `${parse.name} differs under CRLF`);
}
}
});
test('empty, whitespace-only, null and absent input yield empty results, never a throw', () => {
for (const input of ['', ' \n\t\n ', null, undefined, readShipped('docs/does-not-exist.md')]) {
for (const parse of parsers) assert.deepEqual(parse(input), []);
assert.equal(parseProvenanceGrammar(input), null);
assert.equal(parseCannotEnumerateShape(input), null);
assert.deepEqual(parseYamlExampleBlocks(input), []);
assert.equal(sectionContaining(input, 'REQ-UI-03'), '');
}
});
});
describe('#2845 — a back-reference is not a count claim (regression: `next` red on dacae9273)', () => {
// The how-to gained "Dimension 7 is a rule gsd-ui-checker follows, the same as the
// other six dimensions" — correct prose — and the guard read "six dimensions" as that
// document claiming the checker has six. `next` went red on documentation that was
// right.
//
// Worth recording WHY it reached `next`: the docs PR was green. A doc-only diff
// inert-skips the test matrix in the PR lane, so the guard that reads docs never ran
// against the docs change that broke it. It fired on push to `next` — after merge.
test('"the other N dimensions" / "the remaining N dimensions" are not counted', () => {
for (const phrase of [
'Dimension 7 is a rule the checker follows, the same as the other six dimensions.',
'the same as the other 6 dimensions',
'the remaining six dimensions are unchanged',
'the remaining 6 dimensions are unchanged',
// Sentence-initial: the digit scan was case-SENSITIVE while the word scan was not,
// so these two slipped through the first version of this fix.
'Other six dimensions apply.',
'Other 6 dimensions apply.',
'Remaining six dimensions are unaffected.',
'Remaining 6 dimensions are unaffected.',
]) {
assert.deepEqual(parseDeclaredCounts(phrase), [], `counted a back-reference: ${phrase}`);
}
});
test('a real count claim is still counted — the fix must not blind the guard', () => {
assert.deepEqual(parseDeclaredCounts('validates the spec across six dimensions'), [6]);
assert.deepEqual(parseDeclaredCounts('System MUST validate against 7 dimensions'), [7]);
assert.deepEqual(parseDeclaredCounts('All 7 dimensions evaluated'), [7]);
assert.deepEqual(parseDeclaredCounts('**7 Validation Dimensions:**'), [7]);
assert.deepEqual(parseDeclaredCounts('gsd-ui-checker seven quality dimensions'), [7]);
});
test('the exclusion is anchored to the whole word, not a substring', () => {
// "another" contains "other" but has no word boundary before it, so a genuine claim
// about a second set must survive.
assert.deepEqual(parseDeclaredCounts('another six dimensions'), [6]);
});
test('the exclusion is load-bearing on the real shipped how-to, not just on fixtures', () => {
// Typed proof rather than a substring match on prose: run the matcher over the real
// file with the exclusion OFF and then ON. Off, it must report a 6 — that 6 is the
// back-reference, and is literally what turned `next` red. On, only 7s survive.
// If the docs prose is ever reworded away, the first assertion fails loudly rather
// than this guard quietly ceasing to exercise the path it exists for.
const howto = readShipped(SURFACE.HOWTO);
const withoutExclusion = [...new Set(parseDeclaredCounts(howto, { excludeBackReferences: false }))].sort();
const withExclusion = [...new Set(parseDeclaredCounts(howto))].sort();
assert.deepEqual(withoutExclusion, [6, 7],
'vacuous unless the how-to still carries the back-reference this guard exists for');
assert.deepEqual(withExclusion, [7]);
});
});
describe('#2845 — Dimension 7 contract text', () => {
const dim7 = () => sectionContaining(checker, `## Dimension 7: ${DIMENSION_7_LABEL}`);
const dim6 = () => sectionContaining(checker, '## Dimension 6: Registry Safety');
test('declares criteria for all three verdict tiers', () => {
assert.deepEqual([...parseVerdictTiers(dim7())].sort(), ['BLOCK', 'FLAG', 'PASS']);
});
test('ships a well-formed example issue keyed to dimension 7', () => {
const examples = parseYamlExampleBlocks(dim7());
assert.ok(examples.length >= 1, 'Dimension 7 must ship at least one example issue');
const [first] = examples;
assert.equal(first.dimension, '7');
for (const field of ['severity', 'description', 'fix_hint']) {
assert.ok(first[field] && first[field].length > 0, `example issue is missing ${field}`);
}
});
test('records the allowlist downgrade an executor depends on (#2845 acceptance B2)', () => {
const body = lf(dim7()).toLowerCase();
assert.ok(body.includes('closed allowlist'), 'must name what the inventory is NOT');
assert.ok(body.includes('non-exhaustive'), 'must name what it is downgraded TO');
});
test('passes a spec that carries no component inventory at all (backward compatibility)', () => {
const body = lf(dim7()).toLowerCase();
assert.ok(
body.includes('no component inventory'),
'Dimension 7 must state the not-applicable PASS, or every UI-SPEC predating #2845 blocks',
);
});
test('is not scoped by workflow.ui_safety_gate — that clause belongs to Dimension 6 alone', () => {
assert.ok(dim6().includes('workflow.ui_safety_gate'), 'Dimension 6 keeps its config gate');
assert.ok(!dim7().includes('workflow.ui_safety_gate'), 'Dimension 7 must not inherit it');
});
});
describe('#2845 — provenance grammar parity (template emits, checker consumes)', () => {
const templateSlot = () => sectionContaining(template, '## Component Inventory');
const dim7 = () => sectionContaining(checker, `## Dimension 7: ${DIMENSION_7_LABEL}`);
test('the template slot states the grammar and requires all four tokens', () => {
const grammar = parseProvenanceGrammar(templateSlot());
assert.ok(grammar, 'template must state the provenance grammar');
assert.deepEqual([...grammar.tokens].sort(), [...REQUIRED_PROVENANCE_TOKENS].sort());
});
test('the could-not-enumerate shape lives in the SAME slot (#2845 acceptance A3)', () => {
assert.ok(parseCannotEnumerateShape(templateSlot()), 'same slot must accept a negative record');
});
test('the slot states the inventory is non-exhaustive without provenance (#2845 acceptance B2)', () => {
assert.ok(lf(templateSlot()).toLowerCase().includes('non-exhaustive'));
});
test('template and checker quote a byte-identical grammar line', () => {
const emitted = parseProvenanceGrammar(templateSlot());
const consumed = parseProvenanceGrammar(dim7());
assert.ok(emitted && consumed, 'both surfaces must state the grammar');
assert.equal(consumed.line, emitted.line);
assert.deepEqual([...consumed.tokens].sort(), [...emitted.tokens].sort());
});
test('token parity is non-vacuous at 3, 4 and 4-plus-extra tokens', () => {
const four = new Set(REQUIRED_PROVENANCE_TOKENS);
const three = new Set([...four].slice(0, 3)); // limit-1
const extra = new Set([...four, 'unrecognized']); // limit+1
const missing = (set) => REQUIRED_PROVENANCE_TOKENS.filter((t) => !set.has(t));
assert.deepEqual(missing(three), ['date']); // divergence reported
assert.deepEqual(missing(four), []); // exact match
assert.deepEqual(missing(extra), []); // superset still satisfies
});
test('a grammar line missing a token parses as missing it — either surface, both directions', () => {
const full = 'Enumerated by `<command>` — <N> components — <package>@<version> — <YYYY-MM-DD>.';
const noVersion = 'Enumerated by `<command>` — <N> components — <YYYY-MM-DD>.';
const noCommand = 'Enumerated by the design system — <N> components — <package>@<version> — <YYYY-MM-DD>.';
assert.deepEqual([...parseProvenanceGrammar(full).tokens].sort(),
[...REQUIRED_PROVENANCE_TOKENS].sort());
assert.ok(!parseProvenanceGrammar(noVersion).tokens.has(PROVENANCE_TOKEN.VERSION));
assert.ok(!parseProvenanceGrammar(noCommand).tokens.has(PROVENANCE_TOKEN.COMMAND));
assert.notEqual(parseProvenanceGrammar(noVersion).line, parseProvenanceGrammar(full).line);
});
});
describe('#2845 — template structure and researcher duty', () => {
const topHeadings = (text) => lf(text).split('\n')
.map((l) => /^##\s+(\S.*?)\s*$/.exec(l))
.filter(Boolean)
.map((m) => m[1].toLowerCase());
test('the template gains Component Inventory without merging any existing section', () => {
const headings = topHeadings(template);
for (const required of [
'component inventory', 'design system', 'spacing scale', 'typography',
'color', 'copywriting contract', 'ui considerations', 'registry safety',
]) {
assert.ok(headings.includes(required), `template lost or merged "## ${required}"`);
}
});
test('Component Inventory sits with the design system, above the token sections', () => {
const headings = topHeadings(template);
const at = (h) => headings.indexOf(h);
assert.ok(at('design system') < at('component inventory'));
assert.ok(at('component inventory') < at('spacing scale'));
});
test('the researcher is told to enumerate rather than recall, and to record the line', () => {
assert.ok(/enumerat/i.test(lf(researcher)), 'researcher must carry an enumeration duty');
const fromResearcher = parseProvenanceGrammar(researcher);
assert.ok(fromResearcher, 'researcher must quote the provenance grammar');
assert.equal(
fromResearcher.line,
parseProvenanceGrammar(sectionContaining(template, '## Component Inventory')).line,
'researcher and template must quote the identical grammar line',
);
});
});
describe('#2845 — property: roster parity under formatting noise', () => {
const LABELS = [
'Copywriting', 'Visuals', 'Color', 'Typography', 'Spacing',
'Registry Safety', 'Inventory Provenance', 'Motion', 'Density',
'Iconography', 'Localization', 'Elevation',
];
// Document-shaped, not writer-seeded: the arbitrary varies the DOCUMENT (heading
// padding, interleaved unrelated sections, blank-line runs, CRLF) as well as the
// roster, so the property explores shapes a single renderer would never emit (#2371).
const rosterArb = fc.integer({ min: 1, max: LABELS.length })
.map((k) => LABELS.slice(0, k).map((label, i) => ({ n: i + 1, label })));
const noiseArb = fc.record({
crlf: fc.boolean(),
trailingSpaces: fc.boolean(),
interleave: fc.boolean(),
decoy: fc.boolean(),
blankLines: fc.integer({ min: 0, max: 3 }),
});
function renderChecker(roster, noise) {
const pad = noise.trailingSpaces ? ' ' : '';
const gap = '\n'.repeat(noise.blankLines + 1);
const parts = [];
for (const d of roster) {
parts.push(`## Dimension ${d.n}: ${d.label}${pad}`);
parts.push('**Question:** does it hold?');
if (noise.interleave) { parts.push('### Notes'); parts.push('prose'); }
}
// A heading-shaped line inside a fence is an EXAMPLE, never a roster entry. The
// round-trip assertion only holds if the parser skips it, so this decoy is what
// stops the property from being a writer-seeded tautology: a fence-blind parser
// reports roster.length + 1 and the property fails.
if (noise.decoy) {
parts.push('```');
parts.push(`## Dimension ${roster.length + 1}: Decoy`);
parts.push('```');
}
const body = parts.join(gap);
return noise.crlf ? body.replace(/\n/g, '\r\n') : body;
}
function renderVerdict(roster, noise) {
const body = roster.map((d) => `Dimension ${d.n} — ${d.label}: {PASS / FLAG / BLOCK}`).join('\n');
return noise.crlf ? body.replace(/\n/g, '\r\n') : body;
}
test('a rendered roster round-trips, and dropping any one heading always violates', () => {
fc.assert(
fc.property(rosterArb, noiseArb, fc.nat(), (roster, noise, pick) => {
const parsedHeadings = parseDimensionHeadings(renderChecker(roster, noise));
const parsedVerdict = parseVerdictBlock(renderVerdict(roster, noise));
assert.deepEqual(parsedHeadings, roster);
assert.deepEqual(parsedVerdict, roster);
assert.deepEqual(checkRosterParity({
canonical: parsedHeadings,
rosterSurfaces: [{ name: 'verdict', roster: parsedVerdict }],
countSurfaces: [{ name: 'docs', counts: [roster.length] }],
}), []);
// strictly sensitive: remove one dimension from the verdict surface only
const dropped = parsedVerdict.filter((_, i) => i !== pick % roster.length);
assert.deepEqual(
checkRosterParity({
canonical: parsedHeadings,
rosterSurfaces: [{ name: 'verdict', roster: dropped }],
}).map((v) => v.kind),
['roster-mismatch'],
);
}),
{ seed: 2845, numRuns: 200 },
);
});
});