* gsd: Installed
* docs: complete project research
Research for adding GitHub Copilot CLI as 5th runtime to installer.
Files:
- STACK.md: Zero new deps, Copilot reads from .github/, tool name mapping
- FEATURES.md: 18 table stakes, 4 differentiators, 6 anti-features
- ARCHITECTURE.md: Codex-parallel pattern, 5 new functions, 12 existing changes
- PITFALLS.md: 10 pitfalls with prevention strategies and phase mapping
- SUMMARY.md: Synthesized findings, 4-phase roadmap suggestion
* docs(01): create phase plan for core installer plumbing
* feat(01-01): add Copilot as 5th runtime across all install.js locations
- Add --copilot flag parsing and selectedRuntimes integration
- Add 'copilot' to --all array (5 runtimes)
- getDirName('copilot') returns '.github' (local path)
- getGlobalDir('copilot') returns ~/.copilot with COPILOT_CONFIG_DIR override
- getConfigDirFromHome handles copilot for both local/global
- Banner and help text updated to include Copilot
- promptRuntime: Copilot as option 5, All renumbered to option 6
- install(): isCopilot variable, runtimeLabel, skip hooks (Codex pattern)
- install(): Copilot early return before hooks/settings configuration
- finishInstall(): Copilot program name and /gsd-new-project command
- uninstall(): Copilot runtime label and isCopilot variable
- GSD_TEST_MODE exports: getDirName, getGlobalDir, getConfigDirFromHome
* test(01-01): add Copilot plumbing unit tests
- 19 tests covering getDirName, getGlobalDir, getConfigDirFromHome
- getGlobalDir: default path, explicit dir, COPILOT_CONFIG_DIR env var, priority
- Source code integration checks for CLI-01 through CLI-06
- Verifies --both flag unchanged, hooks skipped, prompt options correct
- All 481 tests pass (19 new + 462 existing, no regressions)
* docs(01-01): complete core installer plumbing plan
- Mark Phase 1 and Plan 01-01 as complete in ROADMAP.md
- All 6 requirements (CLI-01 through CLI-06) fulfilled
* gsd: planning
* docs(02): create phase 2 content conversion engine plans
* feat(02-01): add Copilot tool mapping constant and conversion functions
- Add claudeToCopilotTools constant (13 Claude→Copilot tool mappings)
- Add convertCopilotToolName() with mcp__context7__ wildcard handling
- Add convertClaudeToCopilotContent() for CONV-06 (4 path patterns) + CONV-07 (gsd:→gsd-)
- Add convertClaudeCommandToCopilotSkill() for skill frontmatter transformation
- Add convertClaudeAgentToCopilotAgent() with tool dedup and JSON array format
- Export all new functions + constant via GSD_TEST_MODE
* feat(02-01): wire Copilot conversion into install() flow
- Add copyCommandsAsCopilotSkills() for folder-per-skill structure
- Add isCopilot branch in install() skill copy section
- Add isCopilot branch in agent loop with .agent.md rename
- Skip generic path replacement for Copilot (converter handles it)
- Add isCopilot branch in copyWithPathReplacement for .md files
- Add .cjs/.js content transformation for CONV-06/CONV-07
- Export copyCommandsAsCopilotSkills via GSD_TEST_MODE
- CONV-09 not generated (discarded), CONV-10 confirmed working
* docs(02-01): complete content conversion engine plan
- Create 02-01-SUMMARY.md with execution results
- Update STATE.md with Phase 2 position and decisions
- Mark CONV-01 through CONV-10 requirements complete
* test(02-02): add unit tests for Copilot conversion functions
- 16 tests for convertCopilotToolName (all 12 direct mappings, mcp prefix, wildcard, unknown fallback, constant size)
- 8 tests for convertClaudeToCopilotContent (4 path patterns, gsd: conversion, mixed content, no double-replace, passthrough)
- 7 tests for convertClaudeCommandToCopilotSkill (all fields, missing optional fields, CONV-06/07, no frontmatter, agent field)
- 7 tests for convertClaudeAgentToCopilotAgent (dedup, JSON array, field preservation, mcp tools, no tools, CONV-06/07, no frontmatter)
* test(02-02): add integration tests for Copilot skill copy and agent conversion
- copyCommandsAsCopilotSkills produces 31 skill folders with SKILL.md files
- Skill content verified: comma-separated allowed-tools, no YAML multiline, CONV-06/07 applied
- Old skill directories cleaned up on re-run
- gsd-executor agent: 6 tools → 4 after dedup (Write+Edit→edit, Grep+Glob→search)
- gsd-phase-researcher: mcp__context7__* wildcard → io.github.upstash/context7/*
- All 11 agents convert without error, all have frontmatter and tools
- Engine .md and .cjs files: no ~/.claude/ or gsd: references after conversion
- Full suite: 527 tests pass, zero regressions
* docs(02-02): complete Copilot conversion test suite plan
- SUMMARY: 46 new tests covering all conversion functions
- STATE: Phase 02 complete, 3/3 plans done
- ROADMAP: Phase 02 marked complete
* docs(03): research phase domain
* docs(03-instructions-lifecycle): create phase plan
* feat(03-01): add copilot-instructions template and merge/strip functions
- Create get-shit-done/templates/copilot-instructions.md with 5 GSD instructions
- Add GSD_COPILOT_INSTRUCTIONS_MARKER and GSD_COPILOT_INSTRUCTIONS_CLOSE_MARKER constants
- Add mergeCopilotInstructions() with 3-case merge (create, replace, append)
- Add stripGsdFromCopilotInstructions() with null-return for GSD-only content
* feat(03-01): wire install, fix uninstall/manifest/patches for Copilot
- Wire mergeCopilotInstructions() into install() before Copilot early return
- Add else-if isCopilot uninstall branch: remove skills/gsd-*/ + clean instructions
- Fix writeManifest() to hash Copilot skills: (isCodex || isCopilot)
- Fix reportLocalPatches() to show /gsd-reapply-patches for Copilot
- Export new functions and constants in GSD_TEST_MODE
- All 527 existing tests pass with zero regressions
* docs(03-01): complete instructions lifecycle plan
- Create 03-01-SUMMARY.md with execution results
- Update STATE.md: Phase 3 Plan 1 position, decisions, session
- Update ROADMAP.md: Phase 03 progress (1/2 plans)
- Mark INST-01, INST-02, LIFE-01, LIFE-02, LIFE-03 complete
* test(03-02): add unit tests for mergeCopilotInstructions and stripGsdFromCopilotInstructions
- 10 new tests: 5 merge cases + 5 strip cases
- Tests cover create/replace/append merge scenarios
- Tests cover null-return, content preservation, no-markers passthrough
- Added beforeEach/afterEach imports for temp dir lifecycle
- Exported writeManifest and reportLocalPatches via GSD_TEST_MODE for Task 2
* test(03-02): add integration tests for uninstall, manifest, and patches Copilot fixes
- 3 uninstall tests: gsd-* skill identification, instructions cleanup, GSD-only deletion
- writeManifest hashes Copilot skills in manifest JSON (proves isCopilot fix)
- reportLocalPatches uses /gsd-reapply-patches for Copilot (dash format)
- reportLocalPatches uses /gsd:reapply-patches for Claude (no regression)
- Full suite: 543 tests pass, 0 failures
* docs(03-02): complete instructions lifecycle tests plan
- SUMMARY.md with 16 new tests documented
- STATE.md updated: Phase 3 complete, 5/5 plans done
- ROADMAP.md updated: Phase 03 marked complete
* docs(04): capture phase context
* docs(04): research phase domain
* docs(04): create phase plan — E2E integration tests for Copilot install/uninstall
* test(04-01): add E2E Copilot full install verification tests
- 9 tests: skills count/structure, agents count/names, instructions markers
- Manifest structure, categories, SHA256 integrity verification
- Engine directory completeness (bin, references, templates, workflows, CHANGELOG, VERSION)
- Uses execFileSync in isolated /tmp dirs with GSD_TEST_MODE stripped from env
* test(04-01): add E2E Copilot uninstall verification tests
- 6 tests: engine removal, instructions removal, GSD skills/agents cleanup
- Preserves non-GSD custom skills and agents after uninstall
- Standalone lifecycle tests for preservation (install → add custom → uninstall → verify)
- Full suite: 558 tests passing, 0 failures
* docs(04-01): complete E2E Copilot install/uninstall integration tests plan
- SUMMARY.md: 15 E2E tests, SHA256 integrity, 558 total tests passing
- STATE.md: Phase 4 complete, 6/6 plans done
- ROADMAP.md: Phase 4 marked complete
- REQUIREMENTS.md: QUAL-01 complete, QUAL-02 out of scope
* fix: use .github paths for Copilot --local instead of ~/.copilot
convertClaudeToCopilotContent() was hardcoded to always map ~/.claude/
and $HOME/.claude/ to ~/.copilot/ and $HOME/.copilot/ regardless of
install mode. For --local installs these should map to .github/ (repo-
relative, no ./ prefix) since Copilot resolves @file references from
the repo root.
Local mode: ~/.claude/ → .github/ | $HOME/.claude/ → .github/
Global mode: ~/.claude/ → ~/.copilot/ | $HOME/.claude/ → $HOME/.copilot/
Added isGlobal parameter to convertClaudeToCopilotContent,
convertClaudeCommandToCopilotSkill, convertClaudeAgentToCopilotAgent,
copyCommandsAsCopilotSkills, and copyWithPathReplacement. All call
sites in install() now pass isGlobal through.
Tests updated to cover both local (default) and global modes.
565 tests passing, 0 failures.
Co-authored-by: Copilot <223556219+Copilot@users.noreply.github.com>
* fix: use double quotes for argument-hint in Copilot skills
The converter was hardcoding single quotes around argument-hint values
in skill frontmatter. This breaks YAML parsing when the value itself
contains single quotes (e.g., "e.g., 'v1.1 Notifications'").
Now uses yamlQuote() (JSON.stringify) which produces double-quoted
strings with proper escaping.
Co-authored-by: Copilot <223556219+Copilot@users.noreply.github.com>
* chore: complete v1.23 milestone — Copilot CLI Support
Archive milestone artifacts, retrospective, and update project docs.
- Archive: v1.23-ROADMAP.md, v1.23-REQUIREMENTS.md, v1.23-MILESTONE-AUDIT.md
- Create: MILESTONES.md, RETROSPECTIVE.md
- Evolve: PROJECT.md (validated reqs, key decisions, shipped context)
- Reorganize: ROADMAP.md (collapsed v1.23, progress table)
- Update: STATE.md (status: completed)
- Delete: REQUIREMENTS.md (archived, fresh for next milestone)
Co-authored-by: Copilot <223556219+Copilot@users.noreply.github.com>
* chore: archive phase directories from v1.23 milestone
* chore: Clean gsd tracking
* fix: update test counts for new upstream commands and agents
Upstream added validate-phase command (32 skills) and nyquist-auditor agent (12 agents).
Co-authored-by: Copilot <223556219+Copilot@users.noreply.github.com>
* chore: Remove copilot instructions
* chore: Improve loop
* gsd: installation
* docs: start milestone v1.24 Autonomous Skill
* docs: internal research for autonomous skill
* docs: define milestone v1.24 requirements
* docs: create milestone v1.24 roadmap (4 phases)
* docs: phase 5 context — skill scaffolding decisions
* docs(5): research phase domain
* docs(05): create phase plan — 2 plans in 2 waves
* docs(phase-5): add validation strategy
Co-authored-by: Copilot <223556219+Copilot@users.noreply.github.com>
* test(05-01): add failing tests for colon-outside-bold regex format
- Test get-phase with **Goal**: format (colon outside bold)
- Test analyze with **Goal**: and **Depends on**: formats
- Test mixed colon-inside and colon-outside bold formats
- All 3 new tests fail confirming the regex bug
* fix(05-01): fix regex for goal/depends_on extraction in roadmap.cjs
- Fix 3 regex patterns to support both **Goal:** and **Goal**: formats
- Pattern: /\*\*Goal(?::\*\*|\*\*:)/ handles colon inside or outside bold
- Fix in both source (get-shit-done/) and runtime (.github/) copies
- All 28 tests pass including 4 new colon-outside-bold tests
- Live verification: all 4 phases return non-null goals from real ROADMAP.md
* feat(05-01): create gsd:autonomous command file
- name: gsd:autonomous with argument-hint: [--from N]
- Sections: objective, execution_context, context, process
- References workflows/autonomous.md and references/ui-brand.md
- Follows exact pattern of new-milestone.md (42 lines)
* docs(05-01): complete roadmap regex fix + autonomous command plan
* feat(05-02): create autonomous workflow with phase discovery and Skill() execution
- Initialize step with milestone-op bootstrap and --from N flag parsing
- Phase discovery via roadmap analyze with incomplete filtering and sort
- Execute step uses Skill() flat calls for discuss/plan/execute (not Task())
- Progress banner: GSD ► AUTONOMOUS ▸ Phase N/T format with bar
- Iterate step re-reads ROADMAP.md after each phase for dynamic phase detection
- Handle blocker step with retry/skip/stop user options
* test(05-02): add autonomous skill generation tests and fix skill count
- Test autonomous.md converts to gsd-autonomous Copilot skill with correct frontmatter
- Test CONV-07 converts gsd: to gsd- in autonomous command body content
- Update skill count from 32 to 33 (autonomous.md added in plan 01)
- All 645 tests pass across full suite
* docs(05-02): complete autonomous workflow plan
* docs: phase 5 complete — update roadmap and state
* docs: phase 6 context — smart discuss decisions
* docs(06): research smart discuss phase domain
* docs(06): create phase plan
* feat(06-01): replace Skill(discuss-phase) with inline smart discuss
- Add <step name="smart_discuss"> with 5 sub-steps: load prior context, scout codebase, analyze phase with infrastructure detection, present proposals per area in tables, write CONTEXT.md
- Rewire execute_phase step 3a: check has_context before/after, reference smart_discuss inline
- Remove Skill(gsd:discuss-phase) call entirely
- Preserve Skill(gsd:plan-phase) and Skill(gsd:execute-phase) calls unchanged
- Update success criteria to mention smart discuss
- Grey area proposals use table format with recommended/alternative columns
- AskUserQuestion offers Accept all, Change QN, Discuss deeper per area
- Infrastructure phases auto-detected and skip to minimal CONTEXT.md
- CONTEXT.md output uses identical XML-wrapped sections as discuss-phase.md
* docs(06-01): complete smart discuss inline logic plan
* docs(07): phase execution chain context — flag strategy, validation routing, error recovery
* docs(07): research phase execution chain domain
* docs(07): create phase plan
* feat(07-01): wire phase execution chain with verification routing
- Add --no-transition flag to execute-phase Skill() call in step 3c
- Replace step 3d transition with VERIFICATION.md-based routing
- Route on passed/human_needed/gaps_found with appropriate user prompts
- Add gap closure cycle with 1-retry limit to prevent infinite loops
- Route execute-phase failures (no VERIFICATION.md) to handle_blocker
- Update success_criteria with all new verification behaviors
* docs(07-01): complete phase execution chain plan
* docs(08): multi-phase orchestration & lifecycle context
* docs(08): research phase domain
* docs(08): create phase plan
* feat(08-01): add lifecycle step, fix progress bar, document smart_discuss
- Add lifecycle step (audit→complete→cleanup) after all phases complete
- Fix progress bar N/T to use phase number/total milestone phases
- Add smart_discuss CTRL-03 compliance documentation note
- Rewire iterate step to route to lifecycle instead of manual banner
- Renumber handle_blocker from step 5 to step 6
- Add 10 lifecycle-related items to success criteria
- File grows from 630 to 743 lines, 6 to 7 named steps
* docs(08-01): complete multi-phase orchestration & lifecycle plan
* docs: v1.24 milestone audit — passed (18/18 requirements)
* chore: complete v1.24 milestone — Autonomous Skill
* chore: archive phase directories from v1.24 milestone
* gsd: clean
---------
Co-authored-by: Copilot <223556219+Copilot@users.noreply.github.com>
26 KiB
Drive all remaining milestone phases autonomously. For each incomplete phase: discuss → plan → execute using Skill() flat invocations. Pauses only for explicit user decisions (grey area acceptance, blockers, validation requests). Re-reads ROADMAP.md after each phase to catch dynamically inserted phases.
<required_reading>
Read all files referenced by the invoking prompt's execution_context before starting.
</required_reading>
1. Initialize
Parse $ARGUMENTS for --from N flag:
FROM_PHASE=""
if echo "$ARGUMENTS" | grep -qE '\-\-from\s+[0-9]'; then
FROM_PHASE=$(echo "$ARGUMENTS" | grep -oE '\-\-from\s+[0-9]+\.?[0-9]*' | awk '{print $2}')
fi
Bootstrap via milestone-level init:
INIT=$(node "$HOME/.claude/get-shit-done/bin/gsd-tools.cjs" init milestone-op)
Parse JSON for: milestone_version, milestone_name, phase_count, completed_phases, roadmap_exists, state_exists, commit_docs.
If roadmap_exists is false: Error — "No ROADMAP.md found. Run /gsd:new-milestone first."
If state_exists is false: Error — "No STATE.md found. Run /gsd:new-milestone first."
Display startup banner:
━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━
GSD ► AUTONOMOUS
━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━
Milestone: {milestone_version} — {milestone_name}
Phases: {phase_count} total, {completed_phases} complete
If FROM_PHASE is set, display: Starting from phase ${FROM_PHASE}
2. Discover Phases
Run phase discovery:
ROADMAP=$(node "$HOME/.claude/get-shit-done/bin/gsd-tools.cjs" roadmap analyze)
Parse the JSON phases array.
Filter to incomplete phases: Keep only phases where disk_status !== "complete" OR roadmap_complete === false.
Apply --from N filter: If FROM_PHASE was provided, additionally filter out phases where number < FROM_PHASE (use numeric comparison — handles decimal phases like "5.1").
Sort by number in numeric ascending order.
If no incomplete phases remain:
━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━
GSD ► AUTONOMOUS ▸ COMPLETE 🎉
━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━
All phases complete! Nothing left to do.
Exit cleanly.
Display phase plan:
## Phase Plan
| # | Phase | Status |
|---|-------|--------|
| 5 | Skill Scaffolding & Phase Discovery | In Progress |
| 6 | Smart Discuss | Not Started |
| 7 | Auto-Chain Refinements | Not Started |
| 8 | Lifecycle Orchestration | Not Started |
Fetch details for each phase:
DETAIL=$(node "$HOME/.claude/get-shit-done/bin/gsd-tools.cjs" roadmap get-phase ${PHASE_NUM})
Extract phase_name, goal, success_criteria from each. Store for use in execute_phase and transition messages.
3. Execute Phase
For the current phase, display the progress banner:
━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━
GSD ► AUTONOMOUS ▸ Phase {N}/{T}: {Name} [████░░░░] {P}%
━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━
Where N = current phase number (from the ROADMAP, e.g., 6), T = total milestone phases (from phase_count parsed in initialize step, e.g., 8), P = percentage of all milestone phases completed so far. Calculate P as: (number of phases with disk_status "complete" from the latest roadmap analyze / T × 100). Use █ for filled and ░ for empty segments in the progress bar (8 characters wide).
3a. Smart Discuss
Check if CONTEXT.md already exists for this phase:
PHASE_STATE=$(node "$HOME/.claude/get-shit-done/bin/gsd-tools.cjs" init phase-op ${PHASE_NUM})
Parse has_context from JSON.
If has_context is true: Skip discuss — context already gathered. Display:
Phase ${PHASE_NUM}: Context exists — skipping discuss.
Proceed to 3b.
If has_context is false: Execute the smart_discuss step for this phase.
After smart_discuss completes, verify context was written:
PHASE_STATE=$(node "$HOME/.claude/get-shit-done/bin/gsd-tools.cjs" init phase-op ${PHASE_NUM})
Check has_context. If false → go to handle_blocker: "Smart discuss for phase ${PHASE_NUM} did not produce CONTEXT.md."
3b. Plan
Skill(skill="gsd:plan-phase", args="${PHASE_NUM}")
Verify plan produced output — re-run init phase-op and check has_plans. If false → go to handle_blocker: "Plan phase ${PHASE_NUM} did not produce any plans."
3c. Execute
Skill(skill="gsd:execute-phase", args="${PHASE_NUM} --no-transition")
3d. Post-Execution Routing
After execute-phase returns, read the verification result:
VERIFY_STATUS=$(grep "^status:" "${PHASE_DIR}"/*-VERIFICATION.md 2>/dev/null | head -1 | cut -d: -f2 | tr -d ' ')
Where PHASE_DIR comes from the init phase-op call already made in step 3a. If the variable is not in scope, re-fetch:
PHASE_STATE=$(node "$HOME/.claude/get-shit-done/bin/gsd-tools.cjs" init phase-op ${PHASE_NUM})
Parse phase_dir from the JSON.
If VERIFY_STATUS is empty (no VERIFICATION.md or no status field):
Go to handle_blocker: "Execute phase ${PHASE_NUM} did not produce verification results."
If passed:
Display:
Phase ${PHASE_NUM} ✅ ${PHASE_NAME} — Verification passed
Proceed to iterate step.
If human_needed:
Read the human_verification section from VERIFICATION.md to get the count and items requiring manual testing.
Display the items, then ask user via AskUserQuestion:
- question: "Phase ${PHASE_NUM} has items needing manual verification. Validate now or continue to next phase?"
- options: "Validate now" / "Continue without validation"
On "Validate now": Present the specific items from VERIFICATION.md's human_verification section. After user reviews, ask:
- question: "Validation result?"
- options: "All good — continue" / "Found issues"
On "All good — continue": Display Phase ${PHASE_NUM} ✅ Human validation passed and proceed to iterate step.
On "Found issues": Go to handle_blocker with the user's reported issues as the description.
On "Continue without validation": Display Phase ${PHASE_NUM} ⏭ Human validation deferred and proceed to iterate step.
If gaps_found:
Read gap summary from VERIFICATION.md (score and missing items). Display:
⚠ Phase ${PHASE_NUM}: ${PHASE_NAME} — Gaps Found
Score: {N}/{M} must-haves verified
Ask user via AskUserQuestion:
- question: "Gaps found in phase ${PHASE_NUM}. How to proceed?"
- options: "Run gap closure" / "Continue without fixing" / "Stop autonomous mode"
On "Run gap closure": Execute gap closure cycle (limit: 1 attempt):
Skill(skill="gsd:plan-phase", args="${PHASE_NUM} --gaps")
Verify gap plans were created — re-run init phase-op ${PHASE_NUM} and check has_plans. If no new gap plans → go to handle_blocker: "Gap closure planning for phase ${PHASE_NUM} did not produce plans."
Re-execute:
Skill(skill="gsd:execute-phase", args="${PHASE_NUM} --no-transition")
Re-read verification status:
VERIFY_STATUS=$(grep "^status:" "${PHASE_DIR}"/*-VERIFICATION.md 2>/dev/null | head -1 | cut -d: -f2 | tr -d ' ')
If passed or human_needed: Route normally (continue or ask user as above).
If still gaps_found after this retry: Display "Gaps persist after closure attempt." and ask via AskUserQuestion:
- question: "Gap closure did not fully resolve issues. How to proceed?"
- options: "Continue anyway" / "Stop autonomous mode"
On "Continue anyway": Proceed to iterate step. On "Stop autonomous mode": Go to handle_blocker.
This limits gap closure to 1 automatic retry to prevent infinite loops.
On "Continue without fixing": Display Phase ${PHASE_NUM} ⏭ Gaps deferred and proceed to iterate step.
On "Stop autonomous mode": Go to handle_blocker with "User stopped — gaps remain in phase ${PHASE_NUM}".
Smart Discuss
Run smart discuss for the current phase. Proposes grey area answers in batch tables — the user accepts or overrides per area. Produces identical CONTEXT.md output to regular discuss-phase.
Note: Smart discuss is an autonomous-optimized variant of the
gsd:discuss-phaseskill. It produces identical CONTEXT.md output but uses batch table proposals instead of sequential questioning. The originaldiscuss-phaseskill remains unchanged (per CTRL-03). Future milestones may extract this to a separate skill file.
Inputs: PHASE_NUM from execute_phase. Run init to get phase paths:
PHASE_STATE=$(node "$HOME/.claude/get-shit-done/bin/gsd-tools.cjs" init phase-op ${PHASE_NUM})
Parse from JSON: phase_dir, phase_slug, padded_phase, phase_name.
Sub-step 1: Load prior context
Read project-level and prior phase context to avoid re-asking decided questions.
Read project files:
cat .planning/PROJECT.md 2>/dev/null
cat .planning/REQUIREMENTS.md 2>/dev/null
cat .planning/STATE.md 2>/dev/null
Extract from these:
- PROJECT.md — Vision, principles, non-negotiables, user preferences
- REQUIREMENTS.md — Acceptance criteria, constraints, must-haves vs nice-to-haves
- STATE.md — Current progress, decisions logged so far
Read all prior CONTEXT.md files:
find .planning/phases -name "*-CONTEXT.md" 2>/dev/null | sort
For each CONTEXT.md where phase number < current phase:
- Read the
<decisions>section — these are locked preferences - Read
<specifics>— particular references or "I want it like X" moments - Note patterns (e.g., "user consistently prefers minimal UI", "user rejected verbose output")
Build internal prior_decisions context (do not write to file):
<prior_decisions>
## Project-Level
- [Key principle or constraint from PROJECT.md]
- [Requirement affecting this phase from REQUIREMENTS.md]
## From Prior Phases
### Phase N: [Name]
- [Decision relevant to current phase]
- [Preference that establishes a pattern]
</prior_decisions>
If no prior context exists, continue without — expected for early phases.
Sub-step 2: Scout Codebase
Lightweight codebase scan to inform grey area identification and proposals. Keep under ~5% context.
Check for existing codebase maps:
ls .planning/codebase/*.md 2>/dev/null
If codebase maps exist: Read the most relevant ones (CONVENTIONS.md, STRUCTURE.md, STACK.md based on phase type). Extract reusable components, established patterns, integration points. Skip to building context below.
If no codebase maps, do targeted grep:
Extract key terms from the phase goal. Search for related files:
grep -rl "{term1}\|{term2}" src/ app/ --include="*.ts" --include="*.tsx" --include="*.js" --include="*.jsx" 2>/dev/null | head -10
ls src/components/ src/hooks/ src/lib/ src/utils/ 2>/dev/null
Read the 3-5 most relevant files to understand existing patterns.
Build internal codebase_context (do not write to file):
- Reusable assets — existing components, hooks, utilities usable in this phase
- Established patterns — how the codebase does state management, styling, data fetching
- Integration points — where new code connects (routes, nav, providers)
Sub-step 3: Analyze Phase and Generate Proposals
Get phase details:
DETAIL=$(node "$HOME/.claude/get-shit-done/bin/gsd-tools.cjs" roadmap get-phase ${PHASE_NUM})
Extract goal, requirements, success_criteria from the JSON response.
Infrastructure detection — check FIRST before generating grey areas:
A phase is pure infrastructure when ALL of these are true:
- Goal keywords match: "scaffolding", "plumbing", "setup", "configuration", "migration", "refactor", "rename", "restructure", "upgrade", "infrastructure"
- AND success criteria are all technical: "file exists", "test passes", "config valid", "command runs"
- AND no user-facing behavior is described (no "users can", "displays", "shows", "presents")
If infrastructure-only: Skip Sub-step 4. Jump directly to Sub-step 5 with minimal CONTEXT.md. Display:
Phase ${PHASE_NUM}: Infrastructure phase — skipping discuss, writing minimal context.
Use these defaults for the CONTEXT.md:
<domain>: Phase boundary from ROADMAP goal<decisions>: Single "### Claude's Discretion" subsection — "All implementation choices are at Claude's discretion — pure infrastructure phase"<code_context>: Whatever the codebase scout found<specifics>: "No specific requirements — infrastructure phase"<deferred>: "None"
If NOT infrastructure — generate grey area proposals:
Determine domain type from the phase goal:
- Something users SEE → visual: layout, interactions, states, density
- Something users CALL → interface: contracts, responses, errors, auth
- Something users RUN → execution: invocation, output, behavior modes, flags
- Something users READ → content: structure, tone, depth, flow
- Something being ORGANIZED → organization: criteria, grouping, exceptions, naming
Check prior_decisions — skip grey areas already decided in prior phases.
Generate 3-4 grey areas with ~4 questions each. For each question:
- Pre-select a recommended answer based on: prior decisions (consistency), codebase patterns (reuse), domain conventions (standard approaches), ROADMAP success criteria
- Generate 1-2 alternatives per question
- Annotate with prior decision context ("You decided X in Phase N") and code context ("Component Y exists with Z variants") where relevant
Sub-step 4: Present Proposals Per Area
Present grey areas one at a time. For each area (M of N):
Display a table:
### Grey Area {M}/{N}: {Area Name}
| # | Question | ✅ Recommended | Alternative(s) |
|---|----------|---------------|-----------------|
| 1 | {question} | {answer} — {rationale} | {alt1}; {alt2} |
| 2 | {question} | {answer} — {rationale} | {alt1} |
| 3 | {question} | {answer} — {rationale} | {alt1}; {alt2} |
| 4 | {question} | {answer} — {rationale} | {alt1} |
Then prompt the user via AskUserQuestion:
- header: "Area {M}/{N}"
- question: "Accept these answers for {Area Name}?"
- options: Build dynamically — always "Accept all" first, then "Change Q1" through "Change QN" for each question (up to 4), then "Discuss deeper" last. Cap at 6 explicit options max (AskUserQuestion adds "Other" automatically).
On "Accept all": Record all recommended answers for this area. Move to next area.
On "Change QN": Use AskUserQuestion with the alternatives for that specific question:
- header: "{Area Name}"
- question: "Q{N}: {question text}"
- options: List the 1-2 alternatives plus "You decide" (maps to Claude's Discretion)
Record the user's choice. Re-display the updated table with the change reflected. Re-present the full acceptance prompt so the user can make additional changes or accept.
On "Discuss deeper": Switch to interactive mode for this area only — ask questions one at a time using AskUserQuestion with 2-3 concrete options per question plus "You decide". After 4 questions, prompt:
- header: "{Area Name}"
- question: "More questions about {area name}, or move to next?"
- options: "More questions" / "Next area"
If "More questions", ask 4 more. If "Next area", display final summary table of captured answers for this area and move on.
On "Other" (free text): Interpret as either a specific change request or general feedback. Incorporate into the area's decisions, re-display updated table, re-present acceptance prompt.
Scope creep handling: If user mentions something outside the phase domain:
"{Feature} sounds like a new capability — that belongs in its own phase.
I'll note it as a deferred idea.
Back to {current area}: {return to current question}"
Track deferred ideas internally for inclusion in CONTEXT.md.
Sub-step 5: Write CONTEXT.md
After all areas are resolved (or infrastructure skip), write the CONTEXT.md file.
File path: ${phase_dir}/${padded_phase}-CONTEXT.md
Use exactly this structure (identical to discuss-phase output):
# Phase {PHASE_NUM}: {Phase Name} - Context
**Gathered:** {date}
**Status:** Ready for planning
<domain>
## Phase Boundary
{Domain boundary statement from analysis — what this phase delivers}
</domain>
<decisions>
## Implementation Decisions
### {Area 1 Name}
- {Accepted/chosen answer for Q1}
- {Accepted/chosen answer for Q2}
- {Accepted/chosen answer for Q3}
- {Accepted/chosen answer for Q4}
### {Area 2 Name}
- {Accepted/chosen answer for Q1}
- {Accepted/chosen answer for Q2}
...
### Claude's Discretion
{Any "You decide" answers collected — note Claude has flexibility here}
</decisions>
<code_context>
## Existing Code Insights
### Reusable Assets
- {From codebase scout — components, hooks, utilities}
### Established Patterns
- {From codebase scout — state management, styling, data fetching}
### Integration Points
- {From codebase scout — where new code connects}
</code_context>
<specifics>
## Specific Ideas
{Any specific references or "I want it like X" from discussion}
{If none: "No specific requirements — open to standard approaches"}
</specifics>
<deferred>
## Deferred Ideas
{Ideas captured but out of scope for this phase}
{If none: "None — discussion stayed within phase scope"}
</deferred>
Write the file.
Commit:
node "$HOME/.claude/get-shit-done/bin/gsd-tools.cjs" commit "docs(${PADDED_PHASE}): smart discuss context" --files "${phase_dir}/${padded_phase}-CONTEXT.md"
Display confirmation:
Created: {path}
Decisions captured: {count} across {area_count} areas
4. Iterate
After each phase completes, re-read ROADMAP.md to catch phases inserted mid-execution (decimal phases like 5.1):
ROADMAP=$(node "$HOME/.claude/get-shit-done/bin/gsd-tools.cjs" roadmap analyze)
Re-filter incomplete phases using the same logic as discover_phases:
- Keep phases where
disk_status !== "complete"ORroadmap_complete === false - Apply
--from Nfilter if originally provided - Sort by number ascending
Read STATE.md fresh:
cat .planning/STATE.md
Check for blockers in the Blockers/Concerns section. If blockers are found, go to handle_blocker with the blocker description.
If incomplete phases remain: proceed to next phase, loop back to execute_phase.
If all phases complete, proceed to lifecycle step.
5. Lifecycle
After all phases complete, run the milestone lifecycle sequence: audit → complete → cleanup.
Display lifecycle transition banner:
━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━
GSD ► AUTONOMOUS ▸ LIFECYCLE
━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━
All phases complete → Starting lifecycle: audit → complete → cleanup
Milestone: {milestone_version} — {milestone_name}
5a. Audit
Skill(skill="gsd:audit-milestone")
After audit completes, detect the result:
AUDIT_FILE=".planning/v${milestone_version}-MILESTONE-AUDIT.md"
AUDIT_STATUS=$(grep "^status:" "${AUDIT_FILE}" 2>/dev/null | head -1 | cut -d: -f2 | tr -d ' ')
If AUDIT_STATUS is empty (no audit file or no status field):
Go to handle_blocker: "Audit did not produce results — audit file missing or malformed."
If passed:
Display:
Audit ✅ passed — proceeding to complete milestone
Proceed to 5b (no user pause — per CTRL-01).
If gaps_found:
Read the gaps summary from the audit file. Display:
⚠ Audit: Gaps Found
Ask user via AskUserQuestion:
- question: "Milestone audit found gaps. How to proceed?"
- options: "Continue anyway — accept gaps" / "Stop — fix gaps manually"
On "Continue anyway": Display Audit ⏭ Gaps accepted — proceeding to complete milestone and proceed to 5b.
On "Stop": Go to handle_blocker with "User stopped — audit gaps remain. Run /gsd:audit-milestone to review, then /gsd:complete-milestone when ready."
If tech_debt:
Read the tech debt summary from the audit file. Display:
⚠ Audit: Tech Debt Identified
Show the summary, then ask user via AskUserQuestion:
- question: "Milestone audit found tech debt. How to proceed?"
- options: "Continue with tech debt" / "Stop — address debt first"
On "Continue with tech debt": Display Audit ⏭ Tech debt acknowledged — proceeding to complete milestone and proceed to 5b.
On "Stop": Go to handle_blocker with "User stopped — tech debt to address. Run /gsd:audit-milestone to review details."
5b. Complete Milestone
Skill(skill="gsd:complete-milestone", args="${milestone_version}")
After complete-milestone returns, verify it produced output:
ls .planning/milestones/v${milestone_version}-ROADMAP.md 2>/dev/null
If the archive file does not exist, go to handle_blocker: "Complete milestone did not produce expected archive files."
5c. Cleanup
Skill(skill="gsd:cleanup")
Cleanup shows its own dry-run and asks user for approval internally — this is an acceptable pause per CTRL-01 since it's an explicit decision about file deletion.
5d. Final Completion
Display final completion banner:
━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━
GSD ► AUTONOMOUS ▸ COMPLETE 🎉
━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━
Milestone: {milestone_version} — {milestone_name}
Status: Complete ✅
Lifecycle: audit ✅ → complete ✅ → cleanup ✅
Ship it! 🚀
6. Handle Blocker
When any phase operation fails or a blocker is detected, present 3 options via AskUserQuestion:
Prompt: "Phase {N} ({Name}) encountered an issue: {description}"
Options:
- "Fix and retry" — Re-run the failed step (discuss, plan, or execute) for this phase
- "Skip this phase" — Mark phase as skipped, continue to the next incomplete phase
- "Stop autonomous mode" — Display summary of progress so far and exit cleanly
On "Fix and retry": Loop back to the failed step within execute_phase. If the same step fails again after retry, re-present these options.
On "Skip this phase": Log Phase {N} ⏭ {Name} — Skipped by user and proceed to iterate.
On "Stop autonomous mode": Display progress summary:
━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━
GSD ► AUTONOMOUS ▸ STOPPED
━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━
Completed: {list of completed phases}
Skipped: {list of skipped phases}
Remaining: {list of remaining phases}
Resume with: /gsd:autonomous --from {next_phase}
<success_criteria>
- All incomplete phases executed in order (smart discuss → plan → execute each)
- Smart discuss proposes grey area answers in tables, user accepts or overrides per area
- Progress banners displayed between phases
- Execute-phase invoked with --no-transition (autonomous manages transitions)
- Post-execution verification reads VERIFICATION.md and routes on status
- Passed verification → automatic continue to next phase
- Human-needed verification → user prompted to validate or skip
- Gaps-found → user offered gap closure, continue, or stop
- Gap closure limited to 1 retry (prevents infinite loops)
- Plan-phase and execute-phase failures route to handle_blocker
- ROADMAP.md re-read after each phase (catches inserted phases)
- STATE.md checked for blockers before each phase
- Blockers handled via user choice (retry / skip / stop)
- Final completion or stop summary displayed
- After all phases complete, lifecycle step is invoked (not manual suggestion)
- Lifecycle transition banner displayed before audit
- Audit invoked via Skill(skill="gsd:audit-milestone")
- Audit result routing: passed → auto-continue, gaps_found → user decides, tech_debt → user decides
- Audit technical failure (no file/no status) routes to handle_blocker
- Complete-milestone invoked via Skill() with ${milestone_version} arg
- Cleanup invoked via Skill() — internal confirmation is acceptable (CTRL-01)
- Final completion banner displayed after lifecycle
- Progress bar uses phase number / total milestone phases (not position among incomplete)
- Smart discuss documents relationship to discuss-phase with CTRL-03 note </success_criteria>