* fix: Created 10 headless prompt files (5 workflows + 5 agents) in sdk/p… - "sdk/prompts/workflows/execute-plan.md" - "sdk/prompts/workflows/research-phase.md" - "sdk/prompts/workflows/plan-phase.md" - "sdk/prompts/workflows/verify-phase.md" - "sdk/prompts/workflows/discuss-phase.md" - "sdk/prompts/agents/gsd-executor.md" - "sdk/prompts/agents/gsd-phase-researcher.md" - "sdk/prompts/agents/gsd-planner.md" GSD-Task: S01/T02 * feat: Created prompt-sanitizer.ts, wired headless prompt loading into P… - "sdk/src/prompt-sanitizer.ts" - "sdk/src/phase-prompt.ts" - "sdk/src/gsd-tools.ts" - "sdk/src/gsd-tools.test.ts" - "sdk/src/phase-runner-types.test.ts" GSD-Task: S01/T01 * test: Added 111 unit tests covering sanitizePrompt(), headless prompt l… - "sdk/src/prompt-sanitizer.test.ts" - "sdk/src/headless-prompts.test.ts" - "sdk/src/phase-prompt.test.ts" GSD-Task: S01/T03 * feat: Wired sdkPromptsDir preference and sanitizePrompt into InitRunner… - "sdk/src/init-runner.ts" - "sdk/package.json" GSD-Task: S02/T01 * feat: add --init flag to auto command for single-command PRD-to-execution gsd-sdk auto --init @path/to/prd.md now bootstraps the project (init) then immediately runs the autonomous phase execution loop. Co-Authored-By: Claude Opus 4.6 (1M context) <noreply@anthropic.com> * chore: add remaining headless prompt files and templates Co-Authored-By: Claude Opus 4.6 (1M context) <noreply@anthropic.com> * test: Extended init-runner.test.ts with 7 sdkPromptsDir preference and… - "sdk/src/init-runner.test.ts" GSD-Task: S02/T03 --------- Co-authored-by: Claude Opus 4.6 (1M context) <noreply@anthropic.com>
4.3 KiB
4.3 KiB
Execute a phase plan (PLAN.md) and create the outcome summary (SUMMARY.md).
Headless SDK variant — runs autonomously without interactive checkpoints or user prompts.
Load execution context from the session's injected context files. Extract: phase directory, phase number, plans, summaries, incomplete plans, state path, config path.
If verification fails, attempt repair autonomously:
1. Analyze the failure
2. Attempt fix (budget: 2 attempts)
3. If repair succeeds: continue
4. If repair exhausted: log failure, continue with remaining tasks, report in summary
Create SUMMARY.md with:
- Frontmatter: phase, plan, subsystem, tags, dependency graph, tech-stack, key-files, key-decisions, duration, completion timestamp
- Substantive one-liner (not vague)
- Task completion details
- Deviations documentation
- Any blocked items from auth gates or architectural decisions
If planning directory is missing: report error via event stream.
Find the first PLAN without a matching SUMMARY. Decimal phases supported (e.g., `01.1-hotfix/`).Proceed autonomously — no user confirmation needed.
Record plan start timestamp for duration tracking. Check for checkpoint types in the plan:Routing by checkpoint type:
| Checkpoints | Pattern | Execution |
|---|---|---|
| None | A (autonomous) | Execute full plan + SUMMARY |
| Verify-only | B (segmented) | Execute segments autonomously; log verification results instead of pausing |
| Decision | C (main) | Make decisions autonomously based on available context |
In headless mode, all checkpoint types are handled autonomously:
- human-verify checkpoints: run automated verification, log results, continue
- decision checkpoints: select the recommended option (first option), log the choice, continue
- human-action checkpoints: log as a blocker if it requires credentials/auth; otherwise continue with best-effort automation
If plan contains <interfaces> block: Use pre-extracted type definitions directly — do not re-read source files to discover types.
- Read context files from prompt
- Per task:
- MANDATORY read_first gate: If the task has a
<read_first>field, read every listed file BEFORE making edits. type="auto": Implement with deviation rules. Verify done criteria.type="checkpoint:*": Handle autonomously per parse_segments routing above.- MANDATORY acceptance_criteria check: After completing each task, verify EVERY criterion before moving to the next task.
- MANDATORY read_first gate: If the task has a
- Run
<verification>checks - Confirm
<success_criteria>met - Document deviations in Summary
<authentication_gates> Auth errors during execution are interaction points, not failures.
Indicators: "Not authenticated", "Unauthorized", 401/403, "Please run {tool} login", "Set {ENV_VAR}"
Headless protocol:
- Recognize auth gate
- Log the authentication requirement as a blocker event
- Continue with remaining non-blocked tasks
- Report blocked tasks in summary </authentication_gates>
<deviation_rules>
| Rule | Trigger | Action | Permission |
|---|---|---|---|
| 1: Bug | Broken behavior, errors, type errors, security vulns | Fix inline, track [Rule 1 - Bug] |
Auto |
| 2: Missing Critical | Missing error handling, validation, auth, CSRF/CORS | Add inline, track [Rule 2 - Missing Critical] |
Auto |
| 3: Blocking | Prevents completion: missing deps, wrong types, broken imports | Fix blocker, track [Rule 3 - Blocking] |
Auto |
| 4: Architectural | Structural change: new DB table, schema change, new service | Log as blocker event; do NOT proceed with architectural changes autonomously | Report |
| </deviation_rules> |
<success_criteria>
- All tasks from PLAN.md completed (or blocked items documented)
- All verifications pass (or failures documented)
- SUMMARY.md created with substantive content
- Deviations tracked and documented </success_criteria>