* enhance(#4095): checkpoint:decision auto-selection is opt-in via auto_select Auto-mode used to auto-select a checkpoint:decision's first <option> unconditionally, making a decision checkpoint's safety depend on option presentation order rather than an authored choice. Add an optional auto_select="<option-id>" attribute on the <task> tag: absent, auto-mode now escalates to a human exactly like gate="blocking-human" does; present, it names the option auto-mode selects; naming an id with no matching <option id> is a hard structural-validation error at plan-parse time rather than a silent fallback to the first option. gate="blocking-human" continues to win over everything, unchanged. Co-Authored-By: Claude Sonnet 5 <noreply@anthropic.com> * fix(#4095): anchor auto_select/id attribute regexes past hyphenated decoys An isolated adversarial review of the auto_select work found that both new attribute regexes used \b as their left anchor, which is a word boundary, not a "start of attribute name" boundary. A decoy attribute ending in the same word (e.g. data-id="...") sitting before the real id="..." on the same <option> tag matched first, silently corrupting the extracted option id. Anchor on (?:^|\s) instead so only the real attribute name can match. Adds a regression test reproducing the exact decoy-attribute shape, plus a Unicode option-id test from the same review pass. Co-Authored-By: Claude Sonnet 5 <noreply@anthropic.com> * test(#4095): register auto-select-attribute.test.cjs in the docs-guard lane lint-docs-guard-registration failed: the new test reads docs/reference/ plan-md.md but was not registered, so a future edit to that doc could silently desync from the test without the guard catching it on the PR that changed the doc. Registered alongside its direct precedents (precondition-element.test.cjs, reversibility-tagging.test.cjs), which read the same file for the same reason. Co-Authored-By: Claude Sonnet 5 <noreply@anthropic.com> * fix(#4095): fit the decision bullet under execute-phase.md's frozen byte ceiling The remote gsd-test run caught what local checks missed: execute-phase.md carries a frozen ADR-857 Phase-6 byte ceiling (93600) with only 36 bytes of headroom before this change, and the original checkpoint:decision wording pushed it to 93772 (over the ceiling). Cascaded into failures in phase6-capstone-conformance, execute-phase-completion-reconciliation, claude-orchestration, and the compact-content drift-report test. Also caught: tests/package-legitimacy-gate.test.cjs anchors a "decision is conditional, not unconditional" safety check on the literal phrase "first option" in the decision bullet — which #4095 deliberately removes, since there is no more unconditional first-option pick. The test was asserting an assumption this change intentionally makes obsolete; re-anchored on tokens that still identify the bullet ('decision', 'auto-spawn') without weakening what the test actually verifies (the bullet must still carry a blocking-human carve-out). Also fixed a word-order mismatch between my own new test's regex and the actual doc text it was asserting against (tests/auto-select-attribute.test.cjs). Regenerated the compact-content benchmark baseline (tests/fixtures/compact-content-benchmark-baseline.json) to match the new byte counts. Emitted-Drift-Ack-Growth: gsd-executor.md — +7 bytes (49138 -> 49145), from the auto_select carve-out added to the checkpoint:decision auto-mode bullet; already trimmed once to fit the 49152 hard cap. Emitted-Drift-Ack-Growth: execute-phase.md — +12 bytes (93564 -> 93576), from the same carve-out in the orchestrator's decision bullet; kept 24 bytes under the frozen 93600 ADR-857 ceiling after two rounds of trimming for clarity vs. margin. Co-Authored-By: Claude Sonnet 5 <noreply@anthropic.com> * chore(#4095): backfill changeset PR number Co-Authored-By: Claude Sonnet 5 <noreply@anthropic.com> --------- Co-authored-by: sim <sim@local> Co-authored-by: Claude Sonnet 5 <noreply@anthropic.com>
28 KiB
PLAN.md schema reference
A per-plan PLAN.md is GSD Core's executable unit of work — a structured document that tells an executor agent exactly what to build and how to verify it was built correctly. This page documents its structure. See docs index.
Overview
Plans live inside phase directories at:
.planning/phases/<NN>-<slug>/<NN>-<PP>-PLAN.md
For example: .planning/phases/03-post-feed/03-02-PLAN.md (Phase 3, Plan 2).
Plans are produced by the gsd-planner agent (spawned by /gsd-plan-phase) and consumed by execute-phase. A phase typically contains between one and four plans; plans within a phase are assigned to execution waves so that independent work runs in parallel.
YAML frontmatter
Every PLAN.md opens with a YAML frontmatter block between --- delimiters.
Annotated example
---
phase: 03-post-feed
plan: 02
type: execute
wave: 2
depends_on: ["03-01"]
files_modified:
- src/components/PostFeed.tsx
- src/components/PostCard.tsx
- src/app/feed/page.tsx
files_deleted:
- src/components/LegacyFeed.tsx
autonomous: true
requirements: ["FEED-01", "FEED-03"]
user_setup: []
must_haves:
truths:
- "User can scroll through posts from followed accounts"
- "Each post shows author avatar, name, timestamp, and content"
- "Empty state appears when no posts exist"
artifacts:
- path: "src/components/PostFeed.tsx"
provides: "Scrollable post list"
min_lines: 40
- path: "src/components/PostCard.tsx"
provides: "Individual post card"
exports: ["PostCard"]
key_links:
- from: "src/components/PostFeed.tsx"
to: "src/app/api/feed/route.ts"
via: "fetch in useEffect — calls /api/feed endpoint"
pattern: "fetch.*api/feed"
---
Frontmatter field reference
| Field | Required | Type | Purpose |
|---|---|---|---|
phase |
Yes | string | Phase identifier, e.g. 03-post-feed. |
plan |
Yes | string | Plan number within the phase, e.g. 02. |
type |
Yes | execute or tdd |
execute for standard plans; tdd for test-driven plans where tests are written before implementation. |
wave |
Yes | integer | Execution wave. Plans in wave 1 run in parallel (no dependencies). Plans in wave 2+ wait for all plans in the previous wave to complete. Pre-computed at plan time by gsd-planner. |
depends_on |
Yes | array of plan IDs | Plans this plan must wait for. Empty array = wave 1. Accepts three forms, resolved in order: the full plan id ("03-01-auth-hardening"), the canonical phase-plan prefix ("03-01"), or the bare plan number ("01", #3897) — which resolves to the sibling plan in the same phase whose canonical id ends -01. The bare form is in-phase only; it never resolves across phases. If two plans in the same phase share a bare form, the first one (by sorted plan-file order) wins — deterministic, but arbitrary when the collision happens, so prefer the full or canonical form when phase has any short-form collision risk. Example: ["03-01"] means this plan runs after Plan 01 in Phase 3; from within Phase 3 itself, ["01"] means the same thing. |
files_modified |
Yes | array of paths | Every file this plan creates or modifies. Used by the plan-checker to detect same-wave file conflicts and by execute-phase for merge tracking. |
files_deleted |
No | array of paths | Every file this plan deliberately removes. The post-wave cleanup gauntlet blocks the merge of any executor branch whose diff deletes a file — a net against a mass-deletion accident — and this field is the opt-in that names the exceptions. Matching is exact per path after separator normalization: a declared path merges, an undeclared one still blocks that plan's entry (and only that entry). There are no globs and no directory prefixes, so a declaration can never authorize more than it literally lists. Omit the field and the guard's original unconditional block stays in force, which is why absence is always the safe default (#3003). Counts toward same-wave conflict detection alongside files_modified: a plan deleting a file another plan in the same wave is editing is the sharpest conflict there is — one branch removes what the other is writing — so the two plans are pushed into different waves regardless of which side holds the deletion. |
coupling_justified |
No | array of "plan-id: reason" strings |
One entry per deliberately coupled same-wave peer, e.g. ["03-02: both append independent config keys"] — declares that the coupling with that plan through a shared mutable resource (config key, table, migration, env var) is deliberate and order-independent. The plan-checker's Dimension 3b recognizes the declaration and does not flag the pair, so intentionally coupled plans can pass verification without serializing waves. The "plan-id: reason" shape is a prompt-level convention read by the checker, not a schema — the plan parser (src/plan-document.cts) neither validates nor rejects the field, so a typo'd plan-id silently exempts nothing (#3724). |
autonomous |
Yes | boolean | true when all tasks are type auto. false when the plan contains any checkpoint:* task that requires human interaction. |
requirements |
Yes | array of IDs | Requirement IDs from ROADMAP.md that this plan addresses. Every phase requirement ID must appear in at least one plan's requirements field. Empty arrays are a BLOCKER. |
user_setup |
No | array of objects | External-service setup steps that Claude cannot automate (account creation, secret retrieval, dashboard configuration). When present, execute-phase generates a USER-SETUP.md checklist for the developer. |
status |
No | superseded |
Marks a plan that was deliberately reassigned or abandoned mid-phase and will never be executed. A status: superseded plan is excluded from the phase's plan and summary counts, so it never holds the phase below 100%. See Superseded plans. Any other value (or the field's absence) has no effect on counting. |
estimate |
No | object | Projected execution cost: {tokens, raw_tokens, tasks, confidence} (#2631, ADR-2629). tokens is an estimateTokens-scale projection with the project's calibration factor already applied (which is why the plan-checker passes --calibrated to estimate-check — re-applying it would square the correction); confidence (low/med/high) is derived from the calibration sample count, never self-rated. Additive and optional — a plan without it behaves exactly as before. A plan estimated above workflow.smart_zone_tokens is flagged with a split recommendation at plan time; the flag is advisory and never blocks. |
must_haves |
Yes | object | Goal-backward verification criteria. See below. |
agent_hint |
No | string | Per-plan specialist executor routing (#1689). Name of a subagent that shares the gsd-executor execution contract (reads execute-plan.md, atomic-commit protocol). When the named agent resolves on the active runtime (an agent file exists in the runtime's agent dir), execute-phase dispatches it instead of gsd-executor. Unset/unresolved → gsd-executor, byte-identical to today. Default-on via workflow.agent_hint_routing; set false to disable. See Per-plan executor routing. |
gap_closure |
Only in gap-closure mode | string, exact match | Must be exactly the literal lowercase true — validated as a string comparison, not a YAML boolean, so True, TRUE, yes, and 1 are all rejected. Required on every plan generated by /gsd-plan-phase --gaps, checked by the plan-gap-closure schema (src/frontmatter.cts) rather than plan. /gsd-execute-phase --gaps-only filters strictly on this field, so an omitted or wrong-valued gap_closure on a gap-closure plan means it is silently skipped — zero executors spawned, no error (#2847). Standard and reviews-mode plans validate against the unmodified plan schema, which neither requires nor checks this field (nothing rejects it as an extra field either, if present). |
Per-plan executor routing
A plan can opt into a specialist executor by setting agent_hint: to the name of a subagent that shares the gsd-executor execution contract — it reads execute-plan.md, follows the atomic-commit protocol, and carries Read/Edit/Write/Bash. A Flutter specialist, for example:
---
agent_hint: well-me-flutter-engineer
---
At dispatch, execute-phase resolves the hint against the active runtime's agent directory (both project-local and user-global, across the runtime's filename variants — .md, .agent.md, .toml, …) and dispatches the named subagent via subagent_type. If the field is absent, blank, or the named agent does not resolve, the plan dispatches to gsd-executor — byte-identical to behavior without the field. Routing is gated by workflow.agent_hint_routing (default-on; see CONFIGURATION).
The specialist agent is an ordinary agent file (e.g. agents/well-me-flutter-engineer.md on Claude Code); there is no separate registration manifest.
Superseded plans
A phase reads complete when every *-PLAN.md has a matching *-SUMMARY.md. When a plan is reassigned or dropped mid-phase — its work folded into a later plan — it will never gain a summary, and without a marker it would pin the phase below 100% forever (the plan-level analogue of a retired phase). Add status: superseded to that plan's frontmatter to exclude it from both the plan count (denominator) and the summary count (numerator):
---
phase: 05-api
plan: "12"
type: execute
status: superseded
---
A phase with 13 plans, two of them superseded, then reads 11/11 → complete — no fabricated summary required. The match is case-insensitive. Plans without the marker are counted exactly as before.
must_haves field
must_haves captures what must be observably true for the phase goal to be achieved. It is derived during planning and verified after execution by the gsd-verifier agent.
Sub-fields
| Sub-field | Type | Purpose |
|---|---|---|
truths |
array of strings | Observable behaviours from the user's perspective. Each must be verifiable. Example: "User can send a message", not "WebSocket library installed". |
artifacts |
array of objects | Files that must exist with substantive implementation (not stubs). |
artifacts[].path |
string | File path relative to project root. |
artifacts[].provides |
string | What capability this file delivers. |
artifacts[].min_lines |
integer (optional) | Minimum line count to be considered non-stub. |
artifacts[].exports |
array of strings (optional) | Expected named exports to verify. |
artifacts[].contains |
string (optional) | Regex or literal pattern that must appear in the file. |
key_links |
array of objects | Critical connections between artifacts — the wiring that makes the system work end-to-end. |
key_links[].from |
string | Source file (relative path from project root). Must be a literal file path — describe components or symbols in via:. |
key_links[].to |
string | Target file (relative path from project root). Must be a literal file path — describe endpoints, modules, or APIs in via:. |
key_links[].via |
string | Description of how they connect, including any endpoint, component, or symbol name (e.g. fetch in useEffect — calls /api/feed, Prisma query via prisma.message, import). |
key_links[].pattern |
string (optional) | Regex to verify the connection exists in source. |
Body structure
After frontmatter, the plan body uses named XML-style blocks read by the executor agent.
<objective>
States what the plan delivers and why it matters for the project:
<objective>
Implement the post feed as a scrollable card list.
Purpose: Core display feature for the social feed phase.
Output: PostFeed and PostCard components wired to /api/feed.
</objective>
<execution_context>
Lists the workflow files associated with executing the plan. Always includes the execute-plan workflow; adds the checkpoints reference when the plan contains checkpoint tasks:
<execution_context>
@~/.claude/gsd-core/workflows/execute-plan.md
@~/.claude/gsd-core/templates/summary.md
</execution_context>
These @ paths point at the local GSD install, not at repository files. The prefix shown here (~/.claude/gsd-core/…) is the Claude global-install location; other runtimes and local installs resolve to their own install directory — for example .cursor/gsd-core/…, or an absolute project path for a --local install. Because the prefix is install-relative, this block is not clone-portable: a committed plan carries whichever prefix the authoring install had. Execution does not depend on it — /gsd-execute-phase loads the execute-plan workflow from its own installed copy — so the block records the execution context rather than resolvable repository references. Contrast <context> (below), whose repository-relative @ paths resolve after a git clone.
<context>
References source files the executor needs to read. Includes project-level planning docs and any source files whose patterns or types the plan must replicate. Prior plan SUMMARY.md files are included only when there is a genuine dependency (imported types, shared decision) — not reflexively:
<context>
@.planning/PROJECT.md
@.planning/ROADMAP.md
@.planning/STATE.md
@src/components/UserCard.tsx
</context>
<tasks>
Contains one or more <task> elements. Every task element must carry <name>, <files>, <read_first>, <action>, <verify>, <acceptance_criteria>, and <done> for type="auto" and type="tracer" tasks. Optional <precondition> (see Preconditions) and <reversibility> (see Reversibility) elements may sit between <name> and <files>.
Preconditions
<precondition> is an optional element on <task> (issue #1949, The Pragmatic Programmer Topic 23 — Design by Contract). It states, in a single line of runnable/checkable prose, what must already be true for the task to begin safely. It closes the front-of-task side of the contract triad — preconditions (before) ↔ postconditions (<verify>/<done>/<acceptance_criteria>, after) ↔ invariants (must_haves.truths, across the whole plan).
<task type="auto">
<name>Add /reveal endpoint handler</name>
<precondition>server bootstraps and responds to GET /health (from the tracer slice)</precondition>
<files>server/reveal.ts</files>
<action>…</action>
<verify>curl /reveal?path=… opens the OS file manager</verify>
<done>Endpoint committed and manually verified</done>
</task>
Optional and back-compat: a plan that omits <precondition> on every task behaves exactly as today — the executor skips the check with no visible change. Adding <precondition> to a task tells the executor to assert it before any other task work (read-only checks only: file existence, env var presence, idempotent health pings; no side-effecting checks — halt and surface a checkpoint if one seems required) and halt (returning a checkpoint:human-verify, no partial commit) on an unmet precondition. Plans that include <precondition> pass verify plan-structure unchanged — the structural validator checks for the presence of required tags and does not reject unknown optional tags.
Emission cases (planner-side): emit <precondition> only when a task relies on state the plan's own depends_on ordering does not already guarantee. Three cases cover every legitimate use:
- External service setup (
user_setupfrontmatter) — the consuming task ties a specific setup step to itself so the executor halts if the setup was skipped. - Prior-phase artifact dependency — a generated schema, a migration's dist output, a contract file from an earlier phase. Cross-phase
depends_ondoes not cross phase boundaries, so<precondition>is the explicit pointer. - Environment variable / runtime configuration — a tool, API, or script the task invokes requires an env var or runtime config that exists now, not at plan time.
Full emission rules, anti-patterns ("the system is ready" is not checkable; do not use <precondition> for intra-plan sequencing — that is what depends_on is for), and the contract triad mapping: see gsd-core/references/planner-preconditions.md.
Reversibility
<reversibility> is an optional element on <task> (issue #1951, The Pragmatic Programmer Topic 15 — "Reversibility"). It records how costly the decision the task implements would be to undo, so a one-way-door choice gets a human beat before the agent walks through it. The rating attribute carries the classification; the body carries a one-line rationale.
<task type="auto">
<name>Define the on-disk event log format</name>
<reversibility rating="one-way">Phases 4-6 read this file; changing the
format after they land requires a migration for every existing project.</reversibility>
<files>src/event-log.cts</files>
<action>…</action>
<verify><automated>npm run test:unit -- event-log</automated></verify>
<done>Format documented and written by the writer under test</done>
</task>
| Rating | Meaning | Effect on the plan |
|---|---|---|
reversible |
Undo is local and cheap. | None. This is the default when no rating is given. |
costly |
Undo touches many call sites or needs a coordinated change. | Flagged in the plan so the reader sees the weight. Never blocks. |
one-way |
Undo requires a migration, breaks a published contract, or is impossible. | The planner inserts a checkpoint:decision immediately before the dependent task. |
Optional and back-compat: a plan that omits <reversibility> on every task behaves exactly as today — no flag, no checkpoint. Plans that include it pass verify plan-structure unchanged; the structural validator checks for the presence of required tags and does not reject unknown optional tags.
Autonomy: inserting a checkpoint:decision means the plan contains a checkpoint, so its frontmatter must set autonomous: false.
Override: /gsd-plan-phase --no-reversibility-gates (REVERSIBILITY_GATES=false) suppresses checkpoint insertion for intentionally-unattended runs. Ratings are still recorded and costly items are still flagged — the override changes what stops the run, not what the plan remembers.
Full taxonomy, emission rules, and anti-patterns (chiefly: rating everything one-way produces checkpoint fatigue; prefer removing irreversibility over gating it): see gsd-core/references/planner-reversibility.md.
Auto-select
auto_select is an optional attribute on a <task type="checkpoint:decision"> element (issue #4095). It names the id of the <option> that auto-mode (workflow._auto_chain_active / workflow.auto_advance) should select when the checkpoint is reached unattended.
<task type="checkpoint:decision" gate="blocking" auto_select="nextauth">
<decision>Select authentication provider</decision>
<options>
<option id="supabase"><name>Supabase Auth</name></option>
<option id="clerk"><name>Clerk</name></option>
<option id="nextauth"><name>NextAuth.js</name></option>
</options>
<resume-signal>Select: supabase, clerk, or nextauth</resume-signal>
</task>
Semantics:
auto_select |
Auto-mode behavior |
|---|---|
| Absent | Escalates to a human — same treatment as gate="blocking-human". Auto-mode does not guess an answer from option order. |
Names a real <option id="…"> |
Selects that option and logs ⚡ Auto-selected: [option id], then continues. |
Names an id that does not match any <option id="…"> |
verify plan-structure fails at plan-parse time. Never a silent fallback to the first option. |
gate="blocking-human" continues to win over everything, unchanged — it stops for a human in every mode regardless of auto_select.
Why: before this attribute existed, auto-mode always picked the first <option>, and the planner convention of front-loading the recommended choice meant the safety of every decision checkpoint depended on presentation order — a detail no plan author was told was load-bearing. auto_select makes the unattended answer an authored decision instead of a byproduct of layout.
Optional and back-compat for the structural validator: a plan that omits auto_select on every checkpoint:decision task still passes verify plan-structure — the validator only rejects a declared auto_select that doesn't match any option id. What changes is auto-mode's runtime behavior (escalate instead of guessing), not plan-structure validity.
Full behavioral reference: gsd-core/references/checkpoints.md → checkpoint:decision.
Task types
| Type | Use | Autonomy |
|---|---|---|
auto |
Everything the executor can do independently. | Fully autonomous. |
tracer |
The leading thin end-to-end slice a plan starts with by default (tracer-first) — production-quality, wired through every layer, with a real end-to-end <verify>. |
Fully autonomous; after committing, the executor runs the tracer's <verify> as an early integration gate. A tracer carrying gate="blocking-human" STOPs for a human in every mode, auto included. Otherwise autonomous runs halt on failure before expansion, and interactive runs honor workflow.human_verify_mode (#3299): under the end-of-phase default a <verify> carrying only <automated> is re-run and, on success, expansion continues with no checkpoint (failure still halts); under mid-flight, or when the tracer carries <human-check>, a checkpoint:human-verify is presented. Full precedence chain: gsd-core/references/checkpoints.md → "Tracer feedback gate". |
checkpoint:human-verify |
Visual or functional verification that requires a human to look at a running UI or service. | Pauses execution; presents to the developer; resumes on approval. |
checkpoint:decision |
Implementation choices that arose during execution and require the developer's input. | Pauses execution; presents options; resumes on selection. In auto-mode, an auto_select="<option-id>" attribute lets the plan name the unattended answer — see Auto-select. |
checkpoint:human-action |
Truly unavoidable manual steps (account creation, hardware interaction). Used sparingly. | Pauses execution; resumes on confirmation. |
Plans that contain any checkpoint task must set autonomous: false in frontmatter.
auto task structure
<task type="auto">
<name>Task 1: Create PostCard component</name>
<files>src/components/PostCard.tsx</files>
<read_first>src/components/UserCard.tsx, src/types/post.ts</read_first>
<action>Create PostCard component accepting a Post prop (id, authorId, content, createdAt,
reactionCount). Render author avatar using UserAvatar from UserCard pattern. Show timestamp
using date-fns formatDistanceToNow. Export as named export PostCard.</action>
<verify>npx tsc --noEmit</verify>
<acceptance_criteria>
- src/components/PostCard.tsx exports named export PostCard
- PostCard.tsx contains "reactionCount" prop usage
- npx tsc --noEmit exits 0
</acceptance_criteria>
<done>PostCard renders post content with author and timestamp</done>
</task>
Required fields for auto tasks
| Field | Rule |
|---|---|
<files> |
Every file the task creates or modifies. The executor writes only these files. |
<read_first> |
Files the executor must read before touching anything — the file being modified, any source-of-truth pattern file, any file whose types or conventions must be replicated. |
<action> |
Concrete instructions with exact identifiers, file paths, function signatures, and expected values. Never says "align X with Y" without specifying the target state. Never contains fenced code blocks or full implementations. |
<verify> |
A runnable command or check that proves the task succeeded. Must distinguish pass from fail — echo "done" is not valid. Accepts either the wrapped form (<verify><automated>cmd</automated></verify>) or the legacy bare-text form (<verify>cmd</verify>); both are valid. On a type="tracer" task, prefer the wrapped form: the tracer feedback gate's auto-continue (#3299) requires a <verify> carrying only <automated>, so a bare-text tracer verify falls through to the STOP fallback and still presents a checkpoint:human-verify in interactive runs even under end-of-phase. |
<acceptance_criteria> |
Verifiable conditions: grep-verifiable strings, command exit codes, observable behaviours. No subjective language ("looks correct", "properly configured"). Negative greps (! grep -Eq 'PAT' file) are file-scoped — region-scope them (sed -n/awk range, then grep) when a sibling task needs the construct elsewhere in the same file (#968). |
<done> |
A short measurable statement of the completed outcome. |
Plan quality dimensions
The gsd-plan-checker agent reviews every PLAN.md across 12 dimensions before execution begins. A plan that fails any BLOCKER-severity check is returned to gsd-planner for revision (up to 3 iterations):
| Dimension | What it checks |
|---|---|
| 1 — Requirement Coverage | Every phase requirement ID from ROADMAP.md appears in at least one plan's requirements frontmatter field and has covering task(s). |
| 2 — Task Completeness | Every auto task carries all required fields (<files>, <action>, <verify>, <acceptance_criteria>, <done>). No vague or empty fields. |
| 3 — Dependency Correctness | depends_on references are valid, acyclic, and consistent with wave numbers. Wave N plan depends only on plans in waves < N. |
| 4 — Key Links Planned | Artifacts in must_haves.key_links have corresponding tasks that implement the wiring — not just the artifact creation. |
| 5 — Scope Sanity | Plans stay within context budget: 2–3 tasks per plan (4 = warning, 5+ = BLOCKER), ≤ 8–10 files per plan (15+ = BLOCKER). |
| 6 — Verification Derivation | must_haves.truths are user-observable behaviours, not implementation details. Artifacts map to truths. Key links cover critical wiring. |
| 7 — Context Compliance | Every D-NN decision from CONTEXT.md is addressed by at least one task. No task implements anything from <deferred>. |
| 7b — Scope Reduction Detection | Task actions do not silently reduce a locked decision to a "v1", "stub", or "future enhancement" without delivering the full decision scope. Always a BLOCKER when found. |
| 7c — Architectural Tier Compliance | Tasks assign capabilities to the correct tier per the RESEARCH.md Architectural Responsibility Map (when present). Security-sensitive capabilities in the wrong tier are BLOCKERs. |
| 8 — Nyquist Compliance | When workflow.nyquist_validation is enabled and RESEARCH.md exists, every task has an <automated> verify command, no consecutive window of 3 tasks lacks coverage, and VALIDATION.md is present. |
| 9 — Cross-Plan Data Contracts | When plans share data pipelines, their transformations are compatible — no plan strips data that another plan needs in original form. |
| 10 — CLAUDE.md Compliance | Plans respect project-specific conventions, forbidden patterns, required tools, and security requirements from ./CLAUDE.md. |
| 11 — Research Resolution | When RESEARCH.md exists, its ## Open Questions section is marked (RESOLVED) before planning proceeds. |
| 12 — Pattern Compliance | When PATTERNS.md exists, tasks reference the correct analog patterns for each new or modified file. |
Wave execution model
Wave numbers are pre-computed during planning. Execute-phase groups plans by wave number and runs each wave's plans in parallel:
Wave 1: Plan 01, Plan 02, Plan 03 (all run simultaneously — no dependencies)
Wave 2: Plan 04 (waits for Wave 1 to complete)
Wave 3: Plan 05 (waits for Wave 2 to complete)
Plans within a wave that modify overlapping files must not be in the same wave — the plan-checker's Dimension 3 flags this as a BLOCKER.
Plan output
After a plan executes successfully, the executor writes a SUMMARY.md at:
.planning/phases/<NN>-<slug>/<NN>-<PP>-SUMMARY.md
The SUMMARY.md is the canonical record of what was built. Subsequent plans in the same phase may reference it when they have a genuine dependency on its types or decisions.