@mmerterden/multi-agent-pipeline 18.0.0 → 19.1.0
This diff represents the content of publicly available package versions that have been released to one of the supported registries. The information contained in this diff is provided for informational purposes only and reflects changes between package versions as they appear in their respective public registries.
- package/CHANGELOG.md +287 -0
- package/README.md +36 -20
- package/README.tr.md +14 -16
- package/docs/adr/0002-instruction-driven-flag.md +1 -0
- package/docs/adr/0005-lazy-phase-docs.md +11 -1
- package/docs/adr/0008-installer-modularization-and-secret-leak-defense.md +1 -0
- package/docs/adr/0010-own-code-graph.md +1 -0
- package/docs/adr/0014-six-phase-consolidation.md +134 -0
- package/docs/adr/README.md +2 -1
- package/docs/architecture.md +37 -38
- package/docs/best-practices.md +1 -1
- package/docs/ecosystem.md +46 -27
- package/docs/engineering.md +1 -1
- package/docs/facts.json +61 -0
- package/docs/features.md +55 -54
- package/docs/performance.md +5 -5
- package/docs/recovery-guide.md +17 -17
- package/docs/token-budget-history.md +3 -1
- package/index.js +2 -2
- package/install/_codex-agents.mjs +1 -1
- package/install/templates/claude-hooks.json +1 -1
- package/install/templates/codex-instructions.md +1 -1
- package/install/templates/copilot-instructions.md +28 -28
- package/manifest.json +234 -216
- package/package.json +2 -2
- package/pipeline/agents/dev-critic.md +7 -7
- package/pipeline/commands/figma-to-swiftui.md +1 -1
- package/pipeline/commands/multi-agent/SKILL.md +9 -9
- package/pipeline/commands/multi-agent/analysis/SKILL.md +15 -15
- package/pipeline/commands/multi-agent/analysis-resolve/SKILL.md +2 -2
- package/pipeline/commands/multi-agent/autopilot/SKILL.md +7 -7
- package/pipeline/commands/multi-agent/channels/SKILL.md +15 -15
- package/pipeline/commands/multi-agent/diff-explain/SKILL.md +6 -6
- package/pipeline/commands/multi-agent/garbage-collect/SKILL.md +1 -1
- package/pipeline/commands/multi-agent/graph/SKILL.md +1 -1
- package/pipeline/commands/multi-agent/help/SKILL.md +62 -62
- package/pipeline/commands/multi-agent/language/SKILL.md +2 -2
- package/pipeline/commands/multi-agent/local/SKILL.md +11 -11
- package/pipeline/commands/multi-agent/local-autopilot/SKILL.md +13 -13
- package/pipeline/commands/multi-agent/log/SKILL.md +2 -2
- package/pipeline/commands/multi-agent/manual-test/SKILL.md +9 -9
- package/pipeline/commands/multi-agent/model/SKILL.md +69 -0
- package/pipeline/commands/multi-agent/refactor/SKILL.md +3 -3
- package/pipeline/commands/multi-agent/resume/SKILL.md +4 -4
- package/pipeline/commands/multi-agent/resume-local/SKILL.md +19 -17
- package/pipeline/commands/multi-agent/review/SKILL.md +2 -2
- package/pipeline/commands/multi-agent/review-analysis/SKILL.md +1 -1
- package/pipeline/commands/multi-agent/route-off/SKILL.md +36 -0
- package/pipeline/commands/multi-agent/route-on/SKILL.md +74 -0
- package/pipeline/commands/multi-agent/route-status/SKILL.md +56 -0
- package/pipeline/commands/multi-agent/setup/SKILL.md +2 -2
- package/pipeline/commands/multi-agent/status/SKILL.md +5 -5
- package/pipeline/commands/multi-agent/steer/SKILL.md +2 -2
- package/pipeline/commands/multi-agent/sync/SKILL.md +12 -13
- package/pipeline/commands/multi-agent/test/SKILL.md +1 -1
- package/pipeline/lib/credential-inventory.sh +1 -1
- package/pipeline/lib/fetch-fortify.sh +1 -1
- package/pipeline/lib/model-dispatch.sh +140 -0
- package/pipeline/lib/model-rung.sh +142 -0
- package/pipeline/lib/outbound-gate.mjs +14 -0
- package/pipeline/lib/phase-schema.mjs +88 -0
- package/pipeline/lib/plan-todos.sh +5 -5
- package/pipeline/lib/route-state.sh +161 -0
- package/pipeline/lib/run-paths.sh +2 -2
- package/pipeline/multi-agent-refs/_account-picker.md +1 -1
- package/pipeline/multi-agent-refs/_dev-context.md +6 -6
- package/pipeline/multi-agent-refs/_input-parser.md +1 -1
- package/pipeline/multi-agent-refs/analysis/evidence.md +2 -11
- package/pipeline/multi-agent-refs/analysis/intake.md +7 -7
- package/pipeline/multi-agent-refs/analysis/locked.md +48 -22
- package/pipeline/multi-agent-refs/analysis/redesign.md +1 -1
- package/pipeline/multi-agent-refs/analysis/render.md +10 -10
- package/pipeline/multi-agent-refs/analysis/resolve.md +1 -1
- package/pipeline/multi-agent-refs/analysis/review.md +2 -2
- package/pipeline/multi-agent-refs/analysis/synthesis.md +13 -7
- package/pipeline/multi-agent-refs/analysis-template-corporate.md +9 -9
- package/pipeline/multi-agent-refs/analysis-template.md +19 -19
- package/pipeline/multi-agent-refs/android-guide.md +1 -1
- package/pipeline/multi-agent-refs/audit-guide.md +13 -13
- package/pipeline/multi-agent-refs/channels/issue-comment.md +2 -2
- package/pipeline/multi-agent-refs/channels/jira.md +3 -3
- package/pipeline/multi-agent-refs/channels/pr.md +4 -4
- package/pipeline/multi-agent-refs/channels/wiki.md +1 -1
- package/pipeline/multi-agent-refs/component-dispatch.md +8 -8
- package/pipeline/multi-agent-refs/conventions-defaults.md +2 -2
- package/pipeline/multi-agent-refs/cross-cli-contract.md +31 -6
- package/pipeline/multi-agent-refs/features/analysis-jira.md +1 -1
- package/pipeline/multi-agent-refs/features/autopilot-circuit-breaker.md +4 -4
- package/pipeline/multi-agent-refs/features/code-graph.md +5 -5
- package/pipeline/multi-agent-refs/features/design-conformance.md +1 -1
- package/pipeline/multi-agent-refs/features/dev-critic.md +3 -3
- package/pipeline/multi-agent-refs/features/doctor.md +3 -3
- package/pipeline/multi-agent-refs/features/external-context-injection.md +3 -3
- package/pipeline/multi-agent-refs/features/maturity-followup.md +3 -3
- package/pipeline/multi-agent-refs/features/model-fallback.md +41 -5
- package/pipeline/multi-agent-refs/features/plan-todos.md +1 -1
- package/pipeline/multi-agent-refs/features/repo-map.md +1 -1
- package/pipeline/multi-agent-refs/features/review-delta.md +3 -3
- package/pipeline/multi-agent-refs/features/review-multi-repo.md +2 -2
- package/pipeline/multi-agent-refs/features/scope-check.md +4 -4
- package/pipeline/multi-agent-refs/features/skill-conformance.md +2 -2
- package/pipeline/multi-agent-refs/features/stack-skill-routing.md +1 -1
- package/pipeline/multi-agent-refs/features/url-enrichment.md +1 -1
- package/pipeline/multi-agent-refs/features/verify-by-test.md +4 -4
- package/pipeline/multi-agent-refs/features/visual-evidence.md +19 -19
- package/pipeline/multi-agent-refs/features/worktree-finalize.md +6 -6
- package/pipeline/multi-agent-refs/issue-jira-triad.md +10 -10
- package/pipeline/multi-agent-refs/knowledge.md +11 -11
- package/pipeline/multi-agent-refs/multi-repo-integration-build.md +13 -13
- package/pipeline/multi-agent-refs/payload-contracts.md +8 -8
- package/pipeline/multi-agent-refs/phases/log-format.md +10 -10
- package/pipeline/multi-agent-refs/phases/modes.md +30 -30
- package/pipeline/multi-agent-refs/phases/operations.md +8 -8
- package/pipeline/multi-agent-refs/phases/phase-0-init.md +24 -24
- package/pipeline/multi-agent-refs/phases/phase-1-plan.md +599 -0
- package/pipeline/multi-agent-refs/phases/{phase-3-dev.md → phase-2-dev.md} +129 -49
- package/pipeline/multi-agent-refs/phases/{phase-4-review.md → phase-3-review.md} +225 -107
- package/pipeline/multi-agent-refs/phases/{phase-6-commit.md → phase-4-commit.md} +23 -23
- package/pipeline/multi-agent-refs/phases/{phase-7-report.md → phase-5-report.md} +29 -29
- package/pipeline/multi-agent-refs/phases.md +44 -48
- package/pipeline/multi-agent-refs/picker-contract.md +1 -1
- package/pipeline/multi-agent-refs/progress-contract.md +6 -6
- package/pipeline/multi-agent-refs/readiness-review.md +1 -1
- package/pipeline/multi-agent-refs/rules.md +7 -7
- package/pipeline/multi-agent-refs/swiftui-guide.md +2 -2
- package/pipeline/multi-agent-refs/tracker-contract.md +31 -32
- package/pipeline/multi-agent-refs/wiki-capture.md +14 -14
- package/pipeline/preferences-template.json +9 -1
- package/pipeline/rules/figma-pipeline.md +8 -8
- package/pipeline/rules/outside-the-pipeline.md +1 -1
- package/pipeline/schemas/agent-state.schema.json +50 -50
- package/pipeline/schemas/analysis-output.schema.json +3 -3
- package/pipeline/schemas/analysis-spec.schema.json +2 -2
- package/pipeline/schemas/autopilot-config.schema.json +1 -1
- package/pipeline/schemas/code-graph.schema.json +1 -1
- package/pipeline/schemas/criteria-manifest.schema.json +1 -1
- package/pipeline/schemas/dev-critic-output.schema.json +1 -1
- package/pipeline/schemas/diff-risk.schema.json +1 -1
- package/pipeline/schemas/figma-project-config.schema.json +1 -1
- package/pipeline/schemas/migrations/prefs-2.4.0-to-2.5.0.mjs +2 -2
- package/pipeline/schemas/migrations/prefs-2.6.0-to-2.7.0.mjs +31 -0
- package/pipeline/schemas/migrations/state-2.1.0-to-2.2.0.mjs +129 -0
- package/pipeline/schemas/phases.json +105 -0
- package/pipeline/schemas/plan-todos.schema.json +5 -5
- package/pipeline/schemas/planning-output.schema.json +1 -1
- package/pipeline/schemas/prefs.schema.json +102 -58
- package/pipeline/schemas/reviewer-output.schema.json +3 -3
- package/pipeline/schemas/route-config.schema.json +74 -0
- package/pipeline/schemas/scope-check.schema.json +1 -1
- package/pipeline/schemas/secret-patterns.json +124 -0
- package/pipeline/schemas/test-gap.schema.json +1 -1
- package/pipeline/schemas/token-budget.json +12 -18
- package/pipeline/schemas/triage-output.schema.json +6 -6
- package/pipeline/scripts/README.md +3 -3
- package/pipeline/scripts/_code-graph.mjs +2 -2
- package/pipeline/scripts/_run-paths.mjs +2 -2
- package/pipeline/scripts/_smoke-root.sh +1 -1
- package/pipeline/scripts/aggregate-metrics.mjs +1 -1
- package/pipeline/scripts/build-references.mjs +2 -2
- package/pipeline/scripts/bulk-read.sh +10 -1
- package/pipeline/scripts/capture-flush.sh +8 -8
- package/pipeline/scripts/capture-resume.sh +3 -3
- package/pipeline/scripts/classify-plan-safety.mjs +1 -1
- package/pipeline/scripts/cost-table.json +8 -1
- package/pipeline/scripts/diff-explain.mjs +1 -1
- package/pipeline/scripts/doctor.mjs +3 -3
- package/pipeline/scripts/gc-abandoned.sh +3 -3
- package/pipeline/scripts/gc-tmp.sh +1 -1
- package/pipeline/scripts/gc-worktrees.sh +1 -1
- package/pipeline/scripts/gen-facts.mjs +280 -0
- package/pipeline/scripts/gen-mode-dispatch.mjs +32 -37
- package/pipeline/scripts/gen-ref-toc.mjs +1 -1
- package/pipeline/scripts/graph-report.mjs +1 -1
- package/pipeline/scripts/jira-attach.sh +1 -1
- package/pipeline/scripts/learn-from-transcripts.mjs +1 -1
- package/pipeline/scripts/learning-curve.mjs +2 -2
- package/pipeline/scripts/log-metric.sh +17 -4
- package/pipeline/scripts/memory-save.sh +1 -1
- package/pipeline/scripts/migrate-prefs.mjs +22 -5
- package/pipeline/scripts/phase-banner.sh +20 -20
- package/pipeline/scripts/phase-tracker.sh +12 -12
- package/pipeline/scripts/plan-coverage-gate.mjs +2 -2
- package/pipeline/scripts/pre-commit-check.sh +30 -1
- package/pipeline/scripts/render-agent-log-cost.sh +1 -1
- package/pipeline/scripts/render-work-summary.sh +3 -3
- package/pipeline/scripts/review-file-filter.mjs +1 -1
- package/pipeline/scripts/run-aggregator.mjs +13 -6
- package/pipeline/scripts/run-metrics.mjs +1 -1
- package/pipeline/scripts/runs-index.mjs +11 -1
- package/pipeline/scripts/scan-skills.sh +26 -0
- package/pipeline/scripts/smoke-cross-cli-behavior.sh +6 -6
- package/pipeline/scripts/smoke-schema-validation.sh +26 -7
- package/pipeline/scripts/token-budget-report.mjs +13 -2
- package/pipeline/scripts/triage-memory.mjs +2 -2
- package/pipeline/scripts/validate-analysis-doc.mjs +274 -43
- package/pipeline/scripts/validate-planning.mjs +1 -1
- package/pipeline/scripts/validate-reviewer.mjs +1 -1
- package/pipeline/scripts/validate-state.mjs +45 -5
- package/pipeline/scripts/validate-triage.mjs +3 -3
- package/pipeline/scripts/verify-citations.mjs +1 -1
- package/pipeline/scripts/worktree-finalize.sh +5 -5
- package/pipeline/scripts/write-state.mjs +32 -0
- package/pipeline/skills/.skill-manifest.json +38 -22
- package/pipeline/skills/.skills-index.json +49 -5
- package/pipeline/skills/shared/README.md +10 -6
- package/pipeline/skills/shared/core/apple-archive-compliance/SKILL.md +8 -8
- package/pipeline/skills/shared/core/google-play-compliance/SKILL.md +8 -8
- package/pipeline/skills/shared/core/multi-agent/SKILL.md +81 -82
- package/pipeline/skills/shared/core/multi-agent-autopilot/SKILL.md +3 -3
- package/pipeline/skills/shared/core/multi-agent-channels/SKILL.md +14 -14
- package/pipeline/skills/shared/core/multi-agent-diff-explain/SKILL.md +5 -5
- package/pipeline/skills/shared/core/multi-agent-graph/SKILL.md +1 -1
- package/pipeline/skills/shared/core/multi-agent-help/SKILL.md +25 -23
- package/pipeline/skills/shared/core/multi-agent-language/SKILL.md +2 -2
- package/pipeline/skills/shared/core/multi-agent-local/SKILL.md +2 -2
- package/pipeline/skills/shared/core/multi-agent-local-autopilot/SKILL.md +8 -8
- package/pipeline/skills/shared/core/multi-agent-manual-test/SKILL.md +6 -6
- package/pipeline/skills/shared/core/multi-agent-model/SKILL.md +71 -0
- package/pipeline/skills/shared/core/multi-agent-refactor/SKILL.md +3 -3
- package/pipeline/skills/shared/core/multi-agent-resume/SKILL.md +1 -1
- package/pipeline/skills/shared/core/multi-agent-resume-local/SKILL.md +7 -7
- package/pipeline/skills/shared/core/multi-agent-route-off/SKILL.md +39 -0
- package/pipeline/skills/shared/core/multi-agent-route-on/SKILL.md +76 -0
- package/pipeline/skills/shared/core/multi-agent-route-status/SKILL.md +59 -0
- package/pipeline/skills/shared/core/multi-agent-setup/SKILL.md +1 -1
- package/pipeline/skills/shared/core/multi-agent-status/SKILL.md +5 -5
- package/pipeline/skills/shared/core/multi-agent-steer/SKILL.md +2 -2
- package/pipeline/skills/shared/core/multi-agent-sync/SKILL.md +6 -5
- package/pipeline/skills/shared/external/NOTICE-swift-ios-skills.md +1 -1
- package/pipeline/skills/shared/external/signal-community/SKILL.md +8 -1
- package/pipeline/skills/skills-index.md +8 -4
- package/pipeline/multi-agent-refs/phases/phase-1-analysis.md +0 -263
- package/pipeline/multi-agent-refs/phases/phase-2-planning.md +0 -344
- package/pipeline/multi-agent-refs/phases/phase-5-test.md +0 -182
|
@@ -7,7 +7,7 @@
|
|
|
7
7
|
- [Auto-run steps (when host is learned)](#auto-run-steps-when-host-is-learned)
|
|
8
8
|
- [Error evaluation contract](#error-evaluation-contract)
|
|
9
9
|
- [Autopilot behavior](#autopilot-behavior)
|
|
10
|
-
- [Phase
|
|
10
|
+
- [Phase 5 knowledge capture](#phase-5-knowledge-capture)
|
|
11
11
|
- [Generic rule (for copilot-instructions.md)](#generic-rule-for-copilot-instructionsmd)
|
|
12
12
|
- [Schema](#schema)
|
|
13
13
|
- [Smoke coverage](#smoke-coverage)
|
|
@@ -27,7 +27,7 @@ Codegen outputs (identifiers, localization keys, tokens) in Repo A get reference
|
|
|
27
27
|
|
|
28
28
|
## When this rule fires
|
|
29
29
|
|
|
30
|
-
In Phase
|
|
30
|
+
In Phase 4 (Commit & PR), **before** the pre-commit local checkout prompt, if `state.projects.length >= 2`:
|
|
31
31
|
|
|
32
32
|
1. Compute `repoSet` = sorted array of touched repo names.
|
|
33
33
|
2. Look up `prefs.global.multiRepoIntegrationHosts` for an entry whose `repoSet` (sorted) equals this combo.
|
|
@@ -100,7 +100,7 @@ Persist to `prefs.global.multiRepoIntegrationHosts[]`:
|
|
|
100
100
|
|
|
101
101
|
## Auto-run steps (when host is learned)
|
|
102
102
|
|
|
103
|
-
Progress-contract line `→ integrating <host-scheme> with <N> submodules`. Tracker sub-step on Phase
|
|
103
|
+
Progress-contract line `→ integrating <host-scheme> with <N> submodules`. Tracker sub-step on Phase 4 to show live status: `phase-tracker.sh sub 4 0 "Integration build" in_progress`.
|
|
104
104
|
|
|
105
105
|
```bash
|
|
106
106
|
HOST="$(jq -r --arg key "$repoSet_sorted_joined" \
|
|
@@ -126,12 +126,12 @@ BUILD_ERRORS=$(eval "$BUILD_CMD" 2>&1 | grep -E "error:|FAILURE:|error FS" || tr
|
|
|
126
126
|
|
|
127
127
|
# 4. Evaluate + decide
|
|
128
128
|
if [ -z "$BUILD_ERRORS" ]; then
|
|
129
|
-
phase-tracker.sh sub
|
|
130
|
-
log "Phase
|
|
129
|
+
phase-tracker.sh sub 4 0 "Integration build" completed
|
|
130
|
+
log "Phase 4.0: Integration build - clean"
|
|
131
131
|
# proceed to pre-commit checkout prompt + commit
|
|
132
132
|
else
|
|
133
|
-
phase-tracker.sh sub
|
|
134
|
-
log "Phase
|
|
133
|
+
phase-tracker.sh sub 4 0 "Integration build" failed
|
|
134
|
+
log "Phase 4.0: Integration build - NEW errors detected"
|
|
135
135
|
# show errors, ask user
|
|
136
136
|
fi
|
|
137
137
|
```
|
|
@@ -150,13 +150,13 @@ Three outcomes after the build:
|
|
|
150
150
|
Integration build failed with N new errors. Pipeline can go back to
|
|
151
151
|
Phase 3 to fix, or you can fix manually and tell the pipeline to retry.
|
|
152
152
|
|
|
153
|
-
[1] Return to Phase
|
|
153
|
+
[1] Return to Phase 2: Dev with these errors as input (auto-fix attempt)
|
|
154
154
|
[2] Pause - I'll fix manually, then /multi-agent resume
|
|
155
|
-
[3] Override - proceed to commit anyway (logs a warning in Phase
|
|
155
|
+
[3] Override - proceed to commit anyway (logs a warning in Phase 5 report)
|
|
156
156
|
|
|
157
157
|
Select [1-3]:
|
|
158
158
|
```
|
|
159
|
-
3. **Pre-existing errors (unrelated to this task)** → if the pipeline has a baseline error count (from a clean pre-change build), compare deltas. When `current_errors <= baseline`, treat as **no new errors** and proceed; log the pre-existing count in Phase
|
|
159
|
+
3. **Pre-existing errors (unrelated to this task)** → if the pipeline has a baseline error count (from a clean pre-change build), compare deltas. When `current_errors <= baseline`, treat as **no new errors** and proceed; log the pre-existing count in Phase 5 Report.
|
|
160
160
|
|
|
161
161
|
---
|
|
162
162
|
|
|
@@ -164,13 +164,13 @@ Three outcomes after the build:
|
|
|
164
164
|
|
|
165
165
|
Autopilot skips the "Override / pause" prompt. If `count` < 3 on this combo (new / low-confidence), autopilot treats a build failure as BLOCKING and returns to Phase 3 automatically (option 1). If `count >= 3` and `lastResult == success`, a new failure is treated as blocking the same way - but with lower surprise since we've seen this combo succeed before.
|
|
166
166
|
|
|
167
|
-
The learn-once prompt itself also skips under autopilot: instead, autopilot logs `"Phase
|
|
167
|
+
The learn-once prompt itself also skips under autopilot: instead, autopilot logs `"Phase 4.0: SKIPPED - no learned host for this combo, autopilot refuses to prompt. Run in normal mode once to teach."` and proceeds. This is explicit and recoverable.
|
|
168
168
|
|
|
169
169
|
---
|
|
170
170
|
|
|
171
|
-
## Phase
|
|
171
|
+
## Phase 5 knowledge capture
|
|
172
172
|
|
|
173
|
-
When a learn-once prompt writes a new entry, Phase
|
|
173
|
+
When a learn-once prompt writes a new entry, Phase 5 Step 5 (knowledge + memory) ALSO writes a project-scoped memory:
|
|
174
174
|
|
|
175
175
|
```
|
|
176
176
|
Type: reference
|
|
@@ -1,26 +1,26 @@
|
|
|
1
1
|
---
|
|
2
|
-
description: "Canonical required-reading list for outward-facing payloads (PR body, Jira comment, closing report) plus the markup dialect per surface. Loaded by every mode that runs Phase
|
|
2
|
+
description: "Canonical required-reading list for outward-facing payloads (PR body, Jira comment, closing report) plus the markup dialect per surface. Loaded by every mode that runs Phase 4 or Phase 5."
|
|
3
3
|
---
|
|
4
4
|
|
|
5
5
|
# Outward-facing payload contracts
|
|
6
6
|
|
|
7
7
|
> Every mode that opens a PR, comments on a tracker, or closes out a run reads this file first. It does not restate the contracts - it names them, so no mode has to carry its own copy and drift from the others.
|
|
8
8
|
|
|
9
|
-
## Read before Phase
|
|
9
|
+
## Read before Phase 4
|
|
10
10
|
|
|
11
11
|
| Read | Before | Governs |
|
|
12
12
|
|---|---|---|
|
|
13
13
|
| [`channels/pr.md`]($HOME/.claude/multi-agent-refs/channels/pr.md) | assembling the PR body | fixed section set (`summary` → `changes` → `architecture` cond. → `verification` → `risk` cond. → `dependencies` cond. → `related`), Markdown-only rule, reviewer-preserving Bitbucket PUT payload |
|
|
14
|
-
| [`phases/phase-
|
|
14
|
+
| [`phases/phase-4-commit.md`]($HOME/.claude/multi-agent-refs/phases/phase-4-commit.md) | committing | commit convention, default-reviewer fetch, draft/ready prompt, push-must-succeed loop |
|
|
15
15
|
| [`rules.md`]($HOME/.claude/multi-agent-refs/rules.md) "External System Outputs" | any REST payload | real newlines, no HTML entities, no hand-rolled JSON, markup dialect per surface |
|
|
16
16
|
|
|
17
|
-
## Read before Phase
|
|
17
|
+
## Read before Phase 5
|
|
18
18
|
|
|
19
19
|
| Read | Before | Governs |
|
|
20
20
|
|---|---|---|
|
|
21
21
|
| [`channels/jira.md`]($HOME/.claude/multi-agent-refs/channels/jira.md) | posting the Jira comment | fixed section set incl. **Test Scenarios** (Given/When/Then, always present), markdown→wiki conversion table |
|
|
22
22
|
| [`channels/confluence.md`]($HOME/.claude/multi-agent-refs/channels/confluence.md) | writing a Confluence page | storage-format conversion, endpoint flavor |
|
|
23
|
-
| [`phases/phase-
|
|
23
|
+
| [`phases/phase-5-report.md`]($HOME/.claude/multi-agent-refs/phases/phase-5-report.md) | closing out | Timeline + Agent Activity + Cost Breakdown tables, missing-telemetry disclosure |
|
|
24
24
|
| [`tracker-contract.md`]($HOME/.claude/multi-agent-refs/tracker-contract.md) | every phase boundary | per-phase token narration, completion tile suffix |
|
|
25
25
|
|
|
26
26
|
## Markup dialect per surface
|
|
@@ -44,12 +44,12 @@ A payload that is in the right language but the wrong dialect is a defect of the
|
|
|
44
44
|
|
|
45
45
|
The run ends with the phase tracker glyph block **and** the numbers behind it - per-phase duration and token spend, plus totals.
|
|
46
46
|
|
|
47
|
-
- Record spend as you go: `phase-tracker.sh tokens <N> <in> <out> [cached]` after **every** LLM call, including each Phase
|
|
47
|
+
- Record spend as you go: `phase-tracker.sh tokens <N> <in> <out> [cached]` after **every** LLM call, including each Phase 3 reviewer subagent and each Phase 3 chunk. Counts are additive and nothing reconstructs them after the fact.
|
|
48
48
|
- Tag the model once per phase (`phase-tracker.sh model <N> <name>`) or the cost helper cannot price it and prints `-`.
|
|
49
49
|
- Durations come from the phase timestamps and survive a missed `tokens` call; token spend does not.
|
|
50
50
|
- If any phase has no token data, name those phases and say their cost is unavailable. Never print a report whose cost section is simply absent - the reader cannot tell "cheap run" from "nobody recorded it".
|
|
51
51
|
|
|
52
|
-
### Missing-telemetry disclosure (Phase
|
|
52
|
+
### Missing-telemetry disclosure (Phase 5, required)
|
|
53
53
|
|
|
54
54
|
Before composing the report, compute which phases carry no token data:
|
|
55
55
|
|
|
@@ -64,4 +64,4 @@ Non-empty `UNTRACKED` → both the agent-log report and the closing chat summary
|
|
|
64
64
|
|
|
65
65
|
## Fast modes are not exempt
|
|
66
66
|
|
|
67
|
-
A Short run skips Analysis and Planning, and the autopilot and local entries skip the interactive test gate. They run Phase
|
|
67
|
+
A Short run skips Analysis and Planning, and the autopilot and local entries skip the interactive test gate. They run Phase 4 and Phase 5 **unchanged**. A short pipeline is not a licence for an improvised payload shape, a missing Test Scenarios section, or a report without numbers.
|
|
@@ -33,13 +33,13 @@ Create at `$HOME/.claude/logs/multi-agent/{project}/{task-id}/agent-log.md`:
|
|
|
33
33
|
## Phase Duration Distribution
|
|
34
34
|
|
|
35
35
|
Phase 0: Init ████░░░░░░░░░░░░ 5s
|
|
36
|
-
Phase 1:
|
|
37
|
-
Phase
|
|
38
|
-
Phase
|
|
39
|
-
Phase
|
|
40
|
-
Phase
|
|
41
|
-
Phase
|
|
42
|
-
Phase
|
|
36
|
+
Phase 1: Plan (analysis) ██████░░░░░░░░░░ 12s
|
|
37
|
+
Phase 1: Plan █████░░░░░░░░░░░ 8s
|
|
38
|
+
Phase 2: Dev ████████████████ 2m 35s
|
|
39
|
+
Phase 3: Review ████████░░░░░░░░ 37s
|
|
40
|
+
Phase 3: Review (user test) ██░░░░░░░░░░░░░░ (user wait)
|
|
41
|
+
Phase 4: Commit ██░░░░░░░░░░░░░░ 3s
|
|
42
|
+
Phase 5: Report █░░░░░░░░░░░░░░░ 2s
|
|
43
43
|
|
|
44
44
|
## Review Iterations
|
|
45
45
|
|
|
@@ -62,7 +62,7 @@ Verdict: unverified (3 reviewers)
|
|
|
62
62
|
|
|
63
63
|
## Cost Breakdown
|
|
64
64
|
|
|
65
|
-
(emit by Phase
|
|
65
|
+
(emit by Phase 5 via `$HOME/.claude/scripts/render-agent-log-cost.sh <task-id>`. Renders unconditionally on every run. If the renderer exits 2 (no tracker data + no OTel spans), Phase 5 omits this section without failing the run.)
|
|
66
66
|
|
|
67
67
|
| Phase | Model | Tokens in | Tokens out | Est. USD |
|
|
68
68
|
| ----- | ----- | --------- | ---------- | -------- |
|
|
@@ -86,7 +86,7 @@ Verdict: unverified (3 reviewers)
|
|
|
86
86
|
|
|
87
87
|
### Cost Breakdown - emission contract
|
|
88
88
|
|
|
89
|
-
Phase
|
|
89
|
+
Phase 5 MUST attempt to render the Cost Breakdown section as part of the agent-log compose step:
|
|
90
90
|
|
|
91
91
|
```bash
|
|
92
92
|
COST_BLOCK=$(bash $HOME/.claude/scripts/render-agent-log-cost.sh "$TASK_ID" 2>/dev/null) && \
|
|
@@ -97,7 +97,7 @@ Emission is best-effort - exit 2 (no data) is silently skipped. Never fail the
|
|
|
97
97
|
|
|
98
98
|
### Tokens telemetry - phase responsibility
|
|
99
99
|
|
|
100
|
-
Every phase that dispatches a billable LLM agent MUST forward its token totals to the tracker. The minimal contract (already enforced via `smoke-tracker-contract.sh` for Phase
|
|
100
|
+
Every phase that dispatches a billable LLM agent MUST forward its token totals to the tracker. The minimal contract (already enforced via `smoke-tracker-contract.sh` for Phase 3):
|
|
101
101
|
|
|
102
102
|
```bash
|
|
103
103
|
$HOME/.claude/scripts/log-metric.sh "$TASK_ID" <phase-id> <event> \
|
|
@@ -31,27 +31,27 @@ Autopilot mode skips interactive confirmations and runs the pipeline end-to-end
|
|
|
31
31
|
|
|
32
32
|
| Phase | Normal (full) | autopilot (full) | local (full) | local autopilot (full) |
|
|
33
33
|
| ------------------- | ------------------------------------------------------- | ------------------------------------------- | ------------------------------------------- | ------------------------------------------- |
|
|
34
|
-
| Phase
|
|
35
|
-
| Phase
|
|
36
|
-
| Phase
|
|
37
|
-
| Phase
|
|
38
|
-
| **Phase
|
|
34
|
+
| Phase 1 (Plan Approval Gate) | Clarification (max 2 rounds) + approval loop - user: approve/abort/free-text edit | **Gate skip** - log the plan, proceed directly to Phase 2 (autopilot contract: zero interaction) | Same as Normal | Same as autopilot |
|
|
35
|
+
| Phase 3 (User Test) | Interactive prompt ("Want to test?" -> wait) | Skip (autopilot suppresses interactive prompts) -> proceed directly to Phase 4 | **Not in the phase set** (the gate checks the change out of a worktree; local has none) | Not in the phase set |
|
|
36
|
+
| Phase 4 (Commit) | "Want to commit?" -> wait | Auto commit + push | "Want to commit?" -> wait | Auto commit + push |
|
|
37
|
+
| Phase 4 (PR) | "Want to open a PR?" -> wait | Auto create PR | "Want to open a PR?" -> wait | Auto create PR |
|
|
38
|
+
| **Phase 5 (Channels)** | Multi-select channel + content menu | **STILL PAUSES** - see "Phase 5 autopilot exception" below | Multi-select (same as Normal) | **STILL PAUSES** (same as autopilot) |
|
|
39
39
|
|
|
40
40
|
**What NEVER skips (even in autopilot):**
|
|
41
41
|
|
|
42
|
-
- Phase
|
|
42
|
+
- Phase 3 Review -> if blocking finding, returns to Phase 2, auto fix + rebuild (safety)
|
|
43
43
|
- Kill/Purge confirmations -> destructive operations always ask
|
|
44
44
|
- Build fail -> auto fix + rebuild (max 3 retries). After 3 retries still failing -> pause, ask user
|
|
45
45
|
- **Circuit-breaker** -> autopilot halts (records reason, waits for `resume`) on a no-progress stall, an identical repeated failure, a rework storm, cost drift past the `costBudget` ceiling, or a merge/rebase conflict. Full wiring: `$HOME/.claude/multi-agent-refs/features/autopilot-circuit-breaker.md`. This is the sanctioned autopilot pause - continuing unattended off the happy path is the less safe choice.
|
|
46
|
-
- **Phase
|
|
46
|
+
- **Phase 5 channels dispatch** -> always pauses for multi-select. Rationale: Jira/Confluence are externally visible; silently posting wrong-tone content leaks team-visible artifacts. See below.
|
|
47
47
|
|
|
48
48
|
**State tracking**: `agent-state.json` gets `"autopilot": true`. Autopilot continues on resume as well.
|
|
49
49
|
|
|
50
|
-
### Phase
|
|
50
|
+
### Phase 5 autopilot exception
|
|
51
51
|
|
|
52
|
-
The generic "zero-interaction" contract covers Phases 0-
|
|
52
|
+
The generic "zero-interaction" contract covers Phases 0-4 only. Phase 5 channels dispatch is the **single exception**:
|
|
53
53
|
|
|
54
|
-
- ALL modes (`autopilot`, `--local autopilot`, and a Short run reaching Phase
|
|
54
|
+
- ALL modes (`autopilot`, `--local autopilot`, and a Short run reaching Phase 5) pause at the channels multi-select menu.
|
|
55
55
|
- Menu pre-ticks from `prefs.global.reportChannels` + `prefs.global.reportContent` - user can accept with one keypress if prefs are stable.
|
|
56
56
|
- **30-minute timeout** - if user does not respond, session ends cleanly:
|
|
57
57
|
- External delivery aborted (no silent apply - prevents accidental Jira comments / Confluence pages).
|
|
@@ -60,13 +60,13 @@ The generic "zero-interaction" contract covers Phases 0-6 only. Phase 7 channels
|
|
|
60
60
|
- Resume: `/multi-agent:resume <task-id>` re-opens menu with same inputs.
|
|
61
61
|
- Post-hoc `/multi-agent:channels <task>` never times out - user invoked it explicitly.
|
|
62
62
|
|
|
63
|
-
Full contract: `$HOME/.claude/multi-agent-refs/phases/phase-
|
|
63
|
+
Full contract: `$HOME/.claude/multi-agent-refs/phases/phase-5-report.md` (Autopilot pause contract) + `commands/multi-agent/channels/SKILL.md`.
|
|
64
64
|
|
|
65
65
|
---
|
|
66
66
|
|
|
67
67
|
## Pipeline depth (Full / Short)
|
|
68
68
|
|
|
69
|
-
Short strips the pipeline to what an already-scoped task needs: no deep analysis, no planning phase. The **Opus** dev agent reads the task scope and implements, and Phase
|
|
69
|
+
Short strips the pipeline to what an already-scoped task needs: no deep analysis, no planning phase. The **Opus** dev agent reads the task scope and implements, and Phase 3 then reviews what it produced. Review is deliberately NOT part of the strip: analysis and planning shape work that has not happened yet, so a task the user has already scoped can skip them, while review judges work that now exists and has no substitute.
|
|
70
70
|
|
|
71
71
|
**How it is chosen.** Phase 0 Step 7.5, after `taskType` is known:
|
|
72
72
|
|
|
@@ -85,7 +85,7 @@ Bu is icin hangi pipeline? / Which pipeline for this task?
|
|
|
85
85
|
**Pipeline in a Short run:**
|
|
86
86
|
|
|
87
87
|
```
|
|
88
|
-
Phase 0: Init -> Phase
|
|
88
|
+
Phase 0: Init -> Phase 2: Dev (self-contained) -> Phase 3: Review -> Phase 3: Review (user test) -> Phase 4: Commit -> Phase 5: Report
|
|
89
89
|
```
|
|
90
90
|
|
|
91
91
|
**What changes (Tablo 2 - Short runs):**
|
|
@@ -94,14 +94,14 @@ Phase 0: Init -> Phase 3: Dev (self-contained) -> Phase 4: Review -> Phase 5: Te
|
|
|
94
94
|
| ------------------- | ----------------------------------------------- | ---------------------------------------------------------------------------------- | --------------------------------------------- |
|
|
95
95
|
| Phase 0 (Init) | Full setup | Same - worktree, branch, state, and the depth question itself | Same - no worktree, branch on `$PROJECT_ROOT` |
|
|
96
96
|
| Phase 1 (Analysis) | Parallel Explore agents + analysis document | **SKIP** - no tile is ever drawn for it (registration is deferred to Step 7.5) | **SKIP** |
|
|
97
|
-
| Phase
|
|
98
|
-
| Phase
|
|
99
|
-
| Phase
|
|
100
|
-
| Phase
|
|
101
|
-
| Phase
|
|
102
|
-
| Phase
|
|
97
|
+
| Phase 1 (Planning) | TaskCreate + architecture review + **Plan Approval Gate** | **SKIP** (no plan means no plan gate) | **SKIP** |
|
|
98
|
+
| Phase 2 (Dev) | Follows the Phase 1 plan, TDD cycle (Sonnet) | **Self-contained** (Opus): agent scans relevant files, implements with TDD, builds | Same, on the local branch |
|
|
99
|
+
| Phase 3 (Review) | Parallel review + Fable triage (3 reviewers on every host: Claude Code Fable + Opus + Sonnet, Copilot GPT-5.4 + Opus + Sonnet) | **Same** - gates, parallel review, triage; blocking findings return to Phase 2 (cap 3) | **Same**, on the local branch diff |
|
|
100
|
+
| Phase 3 (User Test) | Interactive prompt | **Interactive prompt** | Not in the set - no worktree to check out |
|
|
101
|
+
| Phase 4 (Commit) | Commit + PR | Same - still asks | Same |
|
|
102
|
+
| Phase 5 (Report) | Full report + channels multi-select | Simplified - no analysis section, review section IS present, channels menu still pauses | Same |
|
|
103
103
|
|
|
104
|
-
**Phase
|
|
104
|
+
**Phase 2 in a Short run (self-contained):**
|
|
105
105
|
|
|
106
106
|
The **Opus** agent receives the task description (from Jira, GitHub issue, or free-text) and:
|
|
107
107
|
|
|
@@ -115,7 +115,7 @@ No separate task breakdown - the agent handles scope autonomously.
|
|
|
115
115
|
|
|
116
116
|
### Intake warnings for a Short run
|
|
117
117
|
|
|
118
|
-
**An analysis document was supplied.** Because Phase 1 and Phase
|
|
118
|
+
**An analysis document was supplied.** Because Phase 1 and Phase 1 are skipped, no phase turns that document into a task breakdown. The doc becomes raw context for one Dev pass, and work comes out ordered by whatever the model read first: the bottom of the dependency chain lands, the screen wiring does not.
|
|
119
119
|
|
|
120
120
|
The depth picker is where this is caught. When the intake carried an analysis document or a Figma reference, say so in the question itself rather than after the choice:
|
|
121
121
|
|
|
@@ -134,10 +134,10 @@ Autopilot never sees this: it runs Full.
|
|
|
134
134
|
|
|
135
135
|
## Analysis Mode (`/multi-agent:analysis`)
|
|
136
136
|
|
|
137
|
-
Not a depth permutation - a different **output kind**. It produces the analysis document and stops; no code, no branch, no PR. The
|
|
137
|
+
Not a depth permutation - a different **output kind**. It produces the analysis document and stops; no code, no branch, no PR. The 6-phase contract holds, with four phases reinterpreted the same way a Short run reinterprets Phase 2.
|
|
138
138
|
|
|
139
139
|
```
|
|
140
|
-
Phase 0: Init -> Phase 1:
|
|
140
|
+
Phase 0: Init -> Phase 1: Plan (analysis) -> Phase 1: Plan -> Phase 3: Review -> Phase 4: Publish -> Phase 5: Report
|
|
141
141
|
```
|
|
142
142
|
|
|
143
143
|
| Phase | Analysis mode |
|
|
@@ -153,7 +153,7 @@ Phase 0: Init -> Phase 1: Analysis -> Phase 2: Planning -> Phase 4: Review -> Ph
|
|
|
153
153
|
|
|
154
154
|
**No `local` or `autopilot` variant.** Worktree isolation buys nothing when no code is written, and the intake, the Pass B convention preview and the open-question resolution are interactive by nature; a zero-interaction analysis would be a document nobody agreed to.
|
|
155
155
|
|
|
156
|
-
**Where the document lands.** Drafts go to `/tmp/analysis-<slug>-<ts>/` first, then the Phase
|
|
156
|
+
**Where the document lands.** Drafts go to `/tmp/analysis-<slug>-<ts>/` first, then the Phase 4 picker decides: Local writes `<repo>/analysis/<feature>-<platform>.md` uncommitted, Confluence posts the page, Jira updates the description. In the full pipeline the same engine writes into the worktree instead and `prefs.global.analysisPhase.commitDoc` decides whether it rides along with the commit.
|
|
157
157
|
|
|
158
158
|
---
|
|
159
159
|
|
|
@@ -167,11 +167,11 @@ front so the question resolves without being asked.
|
|
|
167
167
|
```
|
|
168
168
|
Bu is nerede kossun? / Where should this task run?
|
|
169
169
|
1. Worktree .worktrees/{id}/ - your current checkout stays untouched
|
|
170
|
-
2. Lokal / Local the project root, on a new branch - no Phase
|
|
170
|
+
2. Lokal / Local the project root, on a new branch - no Phase 3
|
|
171
171
|
```
|
|
172
172
|
|
|
173
173
|
Two genuine options, so it meets the two-option floor in `picker-contract.md` on
|
|
174
|
-
its own. **Say what local costs inside the question**: Phase
|
|
174
|
+
its own. **Say what local costs inside the question**: Phase 3 is not in a local
|
|
175
175
|
run's set (the user-test gate checks the change out of a worktree, and there is
|
|
176
176
|
none), and uncommitted work in the project root is in the way of the checkout. A
|
|
177
177
|
user choosing local should learn both before choosing, not after.
|
|
@@ -196,9 +196,9 @@ and doing that in the user's own checkout is what worktrees exist to prevent.
|
|
|
196
196
|
| Phase | Normal (worktree) | Local |
|
|
197
197
|
| -------------- | ------------------------------------------------ | -------------------------------------------------------------- |
|
|
198
198
|
| Phase 0 Step 8 | Creates worktree at `.worktrees/{id}/` | `git checkout -b {branch}` directly in `$PROJECT_ROOT` |
|
|
199
|
-
| Phase
|
|
200
|
-
| Phase
|
|
201
|
-
| Phase
|
|
199
|
+
| Phase 2 | Works in worktree path | Works in `$PROJECT_ROOT` |
|
|
200
|
+
| Phase 3 | WIP commit -> remove worktree -> checkout branch | **Not in the phase set** - nothing to check out, nothing to remove |
|
|
201
|
+
| Phase 4 | Commit in worktree, push | Commit in project root, push |
|
|
202
202
|
| All paths | `{worktreePath}` references | All paths use `$PROJECT_ROOT` directly |
|
|
203
203
|
|
|
204
204
|
**Phase 0 Step 8 in local mode:**
|
|
@@ -211,7 +211,7 @@ git -C $PROJECT_ROOT config user.name "{identity.name}"
|
|
|
211
211
|
git -C $PROJECT_ROOT config user.email "{identity.email}"
|
|
212
212
|
```
|
|
213
213
|
|
|
214
|
-
**Phase
|
|
214
|
+
**Phase 3 in local mode:** it is not in the phase set. The gate exists to hand the user a change checked out of a worktree, and in local mode the code is already on the branch in `$PROJECT_ROOT` with nothing to check out or remove. `gen-mode-dispatch.mjs --mode=local` is the enforced statement of this; the row above mirrors it.
|
|
215
215
|
|
|
216
216
|
**State tracking**: `agent-state.json` gets `"localMode": true`, `"worktreePath": null`. All path references resolve to `$PROJECT_ROOT`.
|
|
217
217
|
|
|
@@ -2,7 +2,7 @@
|
|
|
2
2
|
>
|
|
3
3
|
> - **Task IDs** auto-increment from `$HOME/.claude/logs/multi-agent/{project}/.counter` (persistent).
|
|
4
4
|
> - **Kill/Purge/Clear-logs** require explicit user confirm; destructive ops never chain.
|
|
5
|
-
> - **Resume** re-enters at `currentPhase + 1`; always re-runs Phase
|
|
5
|
+
> - **Resume** re-enters at `currentPhase + 1`; always re-runs Phase 5 for knowledge capture.
|
|
6
6
|
> - **3-iteration hard kill**: any retry loop stops after 3 attempts and hands off to the user.
|
|
7
7
|
> - Subagents return JSON (not prose), get only the diff + relevant files, and must reflect before each retry.
|
|
8
8
|
|
|
@@ -61,11 +61,11 @@ Every task gets an auto-incremented short ID. Counter stored at `$HOME/.claude/l
|
|
|
61
61
|
|
|
62
62
|
1. Find state file: `$HOME/.claude/logs/multi-agent/{project}/{task-id}/agent-state.json`
|
|
63
63
|
2. **Validate before re-entry** (required - a half-written or corrupt state silently resumes at the wrong phase):
|
|
64
|
-
- `node $HOME/.claude/scripts/validate-state.mjs <state-file>` (resume-safety check: parseable JSON + `currentPhase` in 0..
|
|
64
|
+
- `node $HOME/.claude/scripts/validate-state.mjs <state-file>` (resume-safety check: parseable JSON + `currentPhase` in 0..5 + well-formed `phases`; tolerant of legacy shapes, not a strict schema match). On non-zero exit, do NOT guess a phase: surface the error and stop with `ERR: agent-state.json is unsafe to resume; inspect it or 'kill #N' and restart.`
|
|
65
65
|
- Confirm the worktree path on `state.worktreePath` / `state.projects[].worktreePath` exists and `git -C <wt> status` is clean-or-known. If the worktree is missing or locked, run the Phase 0 "Worktree stale-lock heal" before continuing.
|
|
66
66
|
3. Read `agent-log.md` for previous findings
|
|
67
67
|
4. Resume from `currentPhase + 1`. If `state.phases[currentPhase+1].subStep` is set, re-enter that phase and skip already-recorded sub-steps (see "Sub-step checkpoints").
|
|
68
|
-
5. **Always run Phase
|
|
68
|
+
5. **Always run Phase 5** on resume completion - ensures knowledge capture even if task was paused mid-pipeline
|
|
69
69
|
6. Log: "Resumed {jiraId} from Phase {N}"
|
|
70
70
|
|
|
71
71
|
---
|
|
@@ -138,13 +138,13 @@ lose the update behind it. `WRITE_STATE_LOCK_STALE_MS` now applies only to a
|
|
|
138
138
|
lock with no readable PID, and `WRITE_STATE_LOCK_ABANDON_MS` (default ten times
|
|
139
139
|
that) is the last-resort ceiling for a PID that has been recycled.
|
|
140
140
|
|
|
141
|
-
**Halt visibility (required, autopilot included).** A halt is never silent. Whenever a phase halts on a hard error (validator failed twice, no subagent returned, dispatch error past fallback, lock irrecoverable), in addition to the `agent-log.md` line: (a) write `state.status = "paused"` and `state.haltReason = "<phase>:<cause>"`; (b) record the cause on the tracker via `phase-tracker.sh meta <phase> halt "<cause>"` and `phase-tracker.sh update <phase> failed`; (c) emit one `>&2` alert line `HALT phase <N>: <cause> - resume with /multi-agent:resume #<id>`; (d) if `prefs.global.usageLog.enabled` is true, emit the end-of-run report so a run that never reaches Phase
|
|
141
|
+
**Halt visibility (required, autopilot included).** A halt is never silent. Whenever a phase halts on a hard error (validator failed twice, no subagent returned, dispatch error past fallback, lock irrecoverable), in addition to the `agent-log.md` line: (a) write `state.status = "paused"` and `state.haltReason = "<phase>:<cause>"`; (b) record the cause on the tracker via `phase-tracker.sh meta <phase> halt "<cause>"` and `phase-tracker.sh update <phase> failed`; (c) emit one `>&2` alert line `HALT phase <N>: <cause> - resume with /multi-agent:resume #<id>`; (d) if `prefs.global.usageLog.enabled` is true, emit the end-of-run report so a run that never reaches Phase 5 is still recorded with the phase it stopped at (`state.currentPhase` + `haltReason`) - the emitter no-ops when it is off or unconfigured:
|
|
142
142
|
|
|
143
143
|
```bash
|
|
144
144
|
node $HOME/.claude/scripts/usage-report.mjs --state "$STATE_FILE" >/dev/null 2>&1 || true
|
|
145
145
|
```
|
|
146
146
|
|
|
147
|
-
Autopilot suppresses *confirmations*, not *halts* - the user must always be able to see why an unattended run stopped without reading the log. The endpoint upserts by run id, so this halt record and a later Phase
|
|
147
|
+
Autopilot suppresses *confirmations*, not *halts* - the user must always be able to see why an unattended run stopped without reading the log. The endpoint upserts by run id, so this halt record and a later Phase 5 record (after resume) collapse into one.
|
|
148
148
|
|
|
149
149
|
### Pipeline Best Practices
|
|
150
150
|
|
|
@@ -160,7 +160,7 @@ This keeps orchestrator context lean and enables programmatic routing.
|
|
|
160
160
|
|
|
161
161
|
**Retry reflection**: Before each retry, force reflection: "What failed? What specific change fixes it? Am I repeating the same approach?" - prevents infinite loops on broken strategies.
|
|
162
162
|
|
|
163
|
-
**Semantic revert**: Track which files each phase produces in `agent-state.json`. If Phase
|
|
163
|
+
**Semantic revert**: Track which files each phase produces in `agent-state.json`. If Phase 3 (testing) fails, revert only Phase 3 (implementation) outputs, not Phase 1 (artifacts):
|
|
164
164
|
|
|
165
165
|
```json
|
|
166
166
|
"phases": {
|
|
@@ -176,7 +176,7 @@ This keeps orchestrator context lean and enables programmatic routing.
|
|
|
176
176
|
bash $HOME/.claude/scripts/capture-flush.sh --state "$STATE_FILE" --quiet
|
|
177
177
|
```
|
|
178
178
|
|
|
179
|
-
The durable stores used to be written only in Phase
|
|
179
|
+
The durable stores used to be written only in Phase 5, which is the phase a run is LEAST likely to reach: a run killed in Phase 2 threw away every finding it had established, and the next run on the same repo paid to rediscover it. The flush is idempotent (the second call through writes 0 rows), costs no API tokens, calls no model, and never fails the transition. Phase 5 is now the LAST flush rather than the only one. The trigger is deliberately mechanical - a phase transition, not a model noticing that a moment qualifies.
|
|
180
180
|
- *Compaction trigger.* If conversation context exceeds ~50%, run `/compact` preserving "modified files, plan, open review findings, current phase + sub-step" before continuing. Don't wait for auto-compaction near the limit - it triggers exactly when context is worst and is lossy. After compaction, re-read `agent-state.json` AND the latest `## Handoff` block in `agent-log.md` to re-ground.
|
|
181
181
|
|
|
182
182
|
**Handoff block (v10.8.0)**: the structured artifact the phase-boundary checkpoint appends to `agent-log.md`. Written by the orchestrator from state it already holds - no agent dispatch, no extra LLM call. Cap at ~15 lines; the latest block is authoritative (earlier ones are history). This is the fresh-context re-entry contract: a resume or post-compaction session rebuilds working context from the latest handoff + `agent-state.json` + git log, never from conversation memory.
|
|
@@ -192,6 +192,6 @@ This keeps orchestrator context lean and enables programmatic routing.
|
|
|
192
192
|
|
|
193
193
|
Full `agent-log.md` shape: `$HOME/.claude/multi-agent-refs/phases/log-format.md`. Resume-side consumption: `resume.md` Step 3 reads the latest handoff FIRST, then falls back to per-phase findings for logs written before v10.8.
|
|
194
194
|
|
|
195
|
-
**Sub-step checkpoints (long phases)**: Phase
|
|
195
|
+
**Sub-step checkpoints (long phases)**: Phase 2 (dev/TDD cycles) and Phase 5 (report/channels) can run many minutes; a crash mid-phase loses everything since the last phase boundary and forces a full phase re-run on resume. For these phases, also record `state.phases[<n>].subStep` (a short token: `red`, `green`, `build`, `pr-opened`, `confluence-synced`, ...) and the `files[]` written so far after each meaningful unit of work. On resume, re-enter the phase but skip units whose `subStep` is already recorded and whose `files[]` exist in the worktree - re-do only the unfinished tail, never the whole phase.
|
|
196
196
|
|
|
197
197
|
**3-iteration hard kill**: Any retry loop (build fix, review fix) MUST stop after 3 attempts. On 4th failure -> pause, ask user. No exceptions.
|
|
@@ -46,7 +46,7 @@ OUTPUT_LANG=$(jq -r '.global.outputLanguage // "en"' "$PREFS_FILE" 2>/dev/null |
|
|
|
46
46
|
|
|
47
47
|
From this point on, everything the user reads renders in `$OUTPUT_LANG`: conversational lines, `AskUserQuestion` `question`/`label`/`description`, and external payload bodies (PR/Jira/Confluence). English stays only on `header`, commit messages, branch names, PR title prefixes, identifiers. Full matrix: `rules.md` "Language Application".
|
|
48
48
|
|
|
49
|
-
**Model tier resolution** (same step, once per run): read `prefs.global.modelFallback`. If `fableEnabled` is `false`, every `preferredModel: fable` persona resolves to `opus` for this run and the Phase
|
|
49
|
+
**Model tier resolution** (same step, once per run): read `prefs.global.modelFallback`. If `fableEnabled` is `false`, every `preferredModel: fable` persona resolves to `opus` for this run and the Phase 3 Claude Code panel is 2 reviewers, not 3; print the one-line INFO. Then, if `premiumTierUntil` is set and in the past, apply the date-gate trigger - `preferredModel` personas dispatch on `fallbackModel`, with the one-line WARN. Both lines and the exact ordering: `$HOME/.claude/multi-agent-refs/features/model-fallback.md`. Dispatch-error and budget triggers apply per-dispatch later; nothing else to do here.
|
|
50
50
|
|
|
51
51
|
**First-run guard**: After loading prefs, check if `keychainMapping` has at least one non-null value. If ALL values are null (template defaults - setup never ran), show:
|
|
52
52
|
```
|
|
@@ -74,7 +74,7 @@ Used for: input parsing, branch naming, commit messages.
|
|
|
74
74
|
|
|
75
75
|
**UX pattern**: Show `Recent: → {value}` suggestion from history, numbered alternatives, enter to accept. No history → skip suggestion line.
|
|
76
76
|
|
|
77
|
-
**Save rule**: After EVERY selection, append chosen value to relevant array (dedup, most recent first, max 10). Save prefs after Phase 0 and Phase
|
|
77
|
+
**Save rule**: After EVERY selection, append chosen value to relevant array (dedup, most recent first, max 10). Save prefs after Phase 0 and Phase 5.
|
|
78
78
|
|
|
79
79
|
**v2.1.0+ Recents update map** (which selection writes to which prefs path):
|
|
80
80
|
|
|
@@ -86,7 +86,7 @@ Used for: input parsing, branch naming, commit messages.
|
|
|
86
86
|
| Service ping (any external API call) | `global.serviceStatus[{service}]` | `{ok, checkedAt, reason?}` (TTL `settings.serviceStatusCacheSeconds`, default 300s) | n/a |
|
|
87
87
|
| Git identity routed (Step 6a) | `projects[{name}].lastIdentity` | int (index into `global.identities`) | n/a |
|
|
88
88
|
|
|
89
|
-
All updates are O(1) - read-modify-write on the in-memory `prefs` object; single atomic write at the end of Phase 0 (and again at Phase
|
|
89
|
+
All updates are O(1) - read-modify-write on the in-memory `prefs` object; single atomic write at the end of Phase 0 (and again at Phase 5).
|
|
90
90
|
|
|
91
91
|
#### Step 0.5 - Figma access pre-flight (BLOCKING when task carries a Figma reference)
|
|
92
92
|
|
|
@@ -96,12 +96,12 @@ Probe order:
|
|
|
96
96
|
|
|
97
97
|
1. **Tier 1 (Figma MCP)**: check the host serves `mcp__claude_ai_Figma__*` before probing. Absent → set `state.figmaAccess.tier1Unavailable = "host"` and fall through to Tier 2 with no probe, no re-auth retry, no MCP-token question. Present → probe `get_metadata(fileKey, nodeId)` on the first frame; on auth failure run `authenticate` + `complete_authentication` and retry once, and only a *second* failure raises the recreate-or-continue question. Success → `state.figmaAccess.tier = 1`.
|
|
98
98
|
2. **Tier 2 (Figma REST)**: when Tier 1 fails, resolve the PAT via `~/.claude/lib/credential-store.sh get <logical-key>` where `<logical-key>` = `prefs.global.keychainMapping.figma`. Probe `GET https://api.figma.com/v1/files/{fileKey}/nodes?ids={nodeId}` with header `X-Figma-Token: $TOKEN`. HTTP 200 → `state.figmaAccess.tier = 2`. Token missing / 401 / 403 → fall through.
|
|
99
|
-
3. **Tier 3 (User-attached screenshot)**: when Tiers 1 + 2 both fail, scan the task payload for inline screenshots or attachments. Present → `state.figmaAccess.tier = 3` and `state.figmaAccess.reviewBlocking = true` (Phase
|
|
99
|
+
3. **Tier 3 (User-attached screenshot)**: when Tiers 1 + 2 both fail, scan the task payload for inline screenshots or attachments. Present → `state.figmaAccess.tier = 3` and `state.figmaAccess.reviewBlocking = true` (Phase 3 enforces this).
|
|
100
100
|
|
|
101
101
|
Save the issue's image attachments to `$WORKTREE/.pipeline/evidence/` as `state.visualEvidence.before[]`: pre-fix evidence, never re-photographed, and none present is a recorded gap rather than a search. See `$HOME/.claude/multi-agent-refs/features/visual-evidence.md`.
|
|
102
102
|
4. **Halt**: all three tiers fail → emit a single AskUserQuestion asking the user how to proceed (provide PAT, paste a screenshot, abort). Never proceed with text-derived guesses.
|
|
103
103
|
|
|
104
|
-
Record the cause, not just the downshift. `tier1Unavailable = "host"` means the tier never existed here - routine on Copilot and Codex, where the installer registers only `multi-agent-toolkit`. `"auth"` means it existed and the credential failed, which on Claude Code points at a dead `figma_mcp` token worth surfacing in Phase
|
|
104
|
+
Record the cause, not just the downshift. `tier1Unavailable = "host"` means the tier never existed here - routine on Copilot and Codex, where the installer registers only `multi-agent-toolkit`. `"auth"` means it existed and the credential failed, which on Claude Code points at a dead `figma_mcp` token worth surfacing in Phase 5. Conflating them costs two wasted MCP round trips and a question the user cannot act on. On those two hosts a mapped `figma` PAT is the primary path, not a fallback.
|
|
105
105
|
|
|
106
106
|
Log the resolved tier in the agent log:
|
|
107
107
|
|
|
@@ -239,13 +239,13 @@ Sequential prompts (standard UX pattern with Recent suggestion): Project Key →
|
|
|
239
239
|
|
|
240
240
|
**Token pre-check** (after parsing): Jira input → resolve key via `prefs.global.keychainMapping.jira`, verify token with a lightweight API call (e.g. `GET /myself`). GitHub input → verify `gh auth status`. On failure (missing key, 401, 403) → run the **Token Save Flow** from `setup.md` inline. This is the same clipboard-based flow used during setup - token never appears in terminal. If user skips and the token is critical for the input type (e.g. Jira token for Jira input), halt Phase 0.
|
|
241
241
|
|
|
242
|
-
**VPN connectivity check**: Test VPN-dependent services (Jira, Bitbucket, Confluence, Fortify, Graylog) with `curl --connect-timeout 3`. If unreachable, warn and offer to continue. Fallback: Jira → manual input, Bitbucket → `git push` only, Confluence → skip Phase
|
|
242
|
+
**VPN connectivity check**: Test VPN-dependent services (Jira, Bitbucket, Confluence, Fortify, Graylog) with `curl --connect-timeout 3`. If unreachable, warn and offer to continue. Fallback: Jira → manual input, Bitbucket → `git push` only, Confluence → skip Phase 5, Fortify → skip scan, Graylog → skip log fetch (advisory only, never blocks). Cache in `agent-state.json` → `"vpnServices": {"jira": true, ...}`.
|
|
243
243
|
|
|
244
244
|
#### Step 1b - URL Enrichment (catalogue + targeted deep fetches)
|
|
245
245
|
|
|
246
246
|
**Runs only when Step 1 found at least one URL in the task input** (Jira/Confluence/Figma/Swagger/Crashlytics/Fortify/Graylog links). Otherwise skip straight to Step 2.
|
|
247
247
|
|
|
248
|
-
When it runs, load `$HOME/.claude/multi-agent-refs/features/url-enrichment.md` and follow it. It covers: 1b.0 extract + catalogue every link into `state.contextLinks[]` (always runs when this step runs), then the deep fetches that are each conditional on their link type being present - 1b.1 Crashlytics -> `state.crashContext`, 1b.2 Fortify SSC -> `state.fortifyFinding`, 1b.3 Graylog, 1b.4 catalogue-only types (fetched at Phase 1). It also owns the two closing log lines and the Phase 1 / Phase
|
|
248
|
+
When it runs, load `$HOME/.claude/multi-agent-refs/features/url-enrichment.md` and follow it. It covers: 1b.0 extract + catalogue every link into `state.contextLinks[]` (always runs when this step runs), then the deep fetches that are each conditional on their link type being present - 1b.1 Crashlytics -> `state.crashContext`, 1b.2 Fortify SSC -> `state.fortifyFinding`, 1b.3 Graylog, 1b.4 catalogue-only types (fetched at Phase 1). It also owns the two closing log lines and the Phase 1 / Phase 1 prepend contracts.
|
|
249
249
|
|
|
250
250
|
#### Step 2 - Project Selection
|
|
251
251
|
|
|
@@ -275,7 +275,7 @@ later because a base branch is a property of a repo set (`picker-contract.md`,
|
|
|
275
275
|
"Order: project, then repo, then branch").
|
|
276
276
|
|
|
277
277
|
Persist `state.siblings[]` even when empty: the empty array is the record that
|
|
278
|
-
the step ran. Absent, Phase
|
|
278
|
+
the step ran. Absent, Phase 3's parity cross-check cannot tell "no siblings"
|
|
279
279
|
from "never asked", and the exit gate below fails.
|
|
280
280
|
|
|
281
281
|
#### Step 3 - Remote Detection + Branch Selection
|
|
@@ -351,7 +351,7 @@ options:
|
|
|
351
351
|
`{BITBUCKET_HOST}` is almost always the VPN - and make the retry real: it re-runs the
|
|
352
352
|
fetch and re-enters this picker on a second failure.
|
|
353
353
|
|
|
354
|
-
Persist the choice in `agent-state.json.baseFetchStatus` ∈ `"fresh" | "cached-stale" | "local-branch" | "aborted"`, and on any non-fresh one log `⚠️ Base ref stale (fetch fail @ {ts}, choice: {status})`. Phase
|
|
354
|
+
Persist the choice in `agent-state.json.baseFetchStatus` ∈ `"fresh" | "cached-stale" | "local-branch" | "aborted"`, and on any non-fresh one log `⚠️ Base ref stale (fetch fail @ {ts}, choice: {status})`. Phase 4 re-attempts `git fetch origin` before push and prompts to rebase if it succeeds.
|
|
355
355
|
|
|
356
356
|
In multi-repo mode the prompt fires per repo, and Abort on any one aborts the whole task (atomic - no partial worktrees).
|
|
357
357
|
|
|
@@ -382,7 +382,7 @@ Branch name is deterministic - no user confirmation needed.
|
|
|
382
382
|
**Collision handling** (automatic - no prompt):
|
|
383
383
|
- Probe local + remote for existing branch. **Distinguish "no such ref" from "the
|
|
384
384
|
probe failed"**: with `2>/dev/null` and an empty-output test a failed probe reads
|
|
385
|
-
as "no collision", and the duplicate branch surfaces as a rejected push at Phase
|
|
385
|
+
as "no collision", and the duplicate branch surfaces as a rejected push at Phase 4.
|
|
386
386
|
```bash
|
|
387
387
|
LOCAL_HIT=$(git -C "$root" rev-parse --verify --quiet "refs/heads/$branch")
|
|
388
388
|
REMOTE_ERR=$(git -C "$root" ls-remote --exit-code --heads origin "$branch" 2>&1 >/dev/null)
|
|
@@ -393,7 +393,7 @@ Branch name is deterministic - no user confirmation needed.
|
|
|
393
393
|
- `REMOTE_RC` is 0 or 2 → treat as authoritative
|
|
394
394
|
- `REMOTE_RC` is anything else → the remote answer is **unknown**, not "free". Log
|
|
395
395
|
`Remote collision probe failed: <REMOTE_ERR>`, fall back to the local check only,
|
|
396
|
-
and record `"remoteCollisionProbe": "failed"` in `agent-state.json` so Phase
|
|
396
|
+
and record `"remoteCollisionProbe": "failed"` in `agent-state.json` so Phase 4
|
|
397
397
|
expects a possible non-fast-forward and re-checks before pushing.
|
|
398
398
|
- No collision → use as-is
|
|
399
399
|
- Collision found → append `-v2`, `-v3`, etc. until unique:
|
|
@@ -548,7 +548,7 @@ Single-repo mode (`projects.length === 1` or scalar-only) uses the legacy single
|
|
|
548
548
|
|
|
549
549
|
**Local-only flow** - when every entry in `state.projects[]` has `provider="local"`:
|
|
550
550
|
- `taskId` format: `LOCAL-{slug-of-freetext}-{yyyymmdd-HHMMSS}` (e.g. `LOCAL-purchase-flow-20260510-143200`). Slug = lowercase, non-alnum → `-`, trimmed, max 32 chars.
|
|
551
|
-
- `state.offlineOnly = true` (Phases
|
|
551
|
+
- `state.offlineOnly = true` (Phases 4/5 read this flag).
|
|
552
552
|
- `state.remoteType` per project = `"local"`.
|
|
553
553
|
- `state.baseBranch` = current branch of the local checkout (no `origin/{base}` fetch).
|
|
554
554
|
- `state.branch` = local-only feature branch on the same checkout; no upstream tracking is configured (`git checkout -b {branch}` without `-u`).
|
|
@@ -577,10 +577,10 @@ Persist: `"taskType": "component" | "bugfix" | "feature" | "refactor" | "chore"`
|
|
|
577
577
|
|
|
578
578
|
| Phase | Behavior change |
|
|
579
579
|
| ------- | -------------------------------------------------------------------------------------------------------- |
|
|
580
|
-
| Phase
|
|
581
|
-
| Phase
|
|
582
|
-
| Phase
|
|
583
|
-
| Phase
|
|
580
|
+
| Phase 2 | `component` → dispatch to the enabled `ai-<platform>-toolkit` plugin's `create-component` skill (fallback `create-ui-component`); else standard TDD |
|
|
581
|
+
| Phase 3 | `bugfix` → test coverage; `component` → accessibility+tokens; `refactor` → behavior preservation |
|
|
582
|
+
| Phase 4 | `bugfix`/`hotfix` → `fix(...)` prefix; `feature`/`component` → `feat(...)`; `refactor` → `refactor(...)` |
|
|
583
|
+
| Phase 5 | `component` → includes SubPhase breakdown |
|
|
584
584
|
|
|
585
585
|
Log: `Phase 0 Step 7: taskType = {component|bugfix|feature|refactor|chore}`
|
|
586
586
|
|
|
@@ -610,7 +610,7 @@ Log: `Phase 0 Step 7.5: depth = {full|short} (recommended {full|short}, source {
|
|
|
610
610
|
|
|
611
611
|
#### Step 7.6 - Test baseline (opt-in, `prefs.global.testBaseline.enabled`, default `false`)
|
|
612
612
|
|
|
613
|
-
Phase
|
|
613
|
+
Phase 3 Gate 3 cannot tell an inherited red suite from one this run broke, so it blocks on someone else's bug or the agent "fixes" tests it never touched. Runs after Step 6, only when the stack has a test command; skipped in analysis mode.
|
|
614
614
|
|
|
615
615
|
```bash
|
|
616
616
|
BASELINE_LOG="$WORKTREE/.baseline-test.log"
|
|
@@ -625,7 +625,7 @@ Log: `Phase 0 Step 7.6: test baseline = {green|red|unknown} ({N} pre-existing fa
|
|
|
625
625
|
|
|
626
626
|
Decide, then probe, then ask, then run. Skipped only when `visualEvidence.enabled` is `false`. Contract: `features/visual-evidence.md` sections 1a, 1b, 4.
|
|
627
627
|
|
|
628
|
-
**This step is the only writer of `visualEvidence.required`, `requiredBy` and `platform`** (section 1a). No diff exists yet, so the `bugfix` row is provisional and Phase
|
|
628
|
+
**This step is the only writer of `visualEvidence.required`, `requiredBy` and `platform`** (section 1a). No diff exists yet, so the `bugfix` row is provisional and Phase 2 re-decides.
|
|
629
629
|
|
|
630
630
|
```bash
|
|
631
631
|
eval "$(bash $HOME/.claude/lib/stack-detect.sh "$PROJECT_ROOT")"
|
|
@@ -653,7 +653,7 @@ DEPTH_DEFAULT_INDEX=1
|
|
|
653
653
|
ASK_CHOICE_DEFAULT="$DEPTH_DEFAULT_INDEX" $HOME/.claude/lib/ask-choice.sh ...
|
|
654
654
|
```
|
|
655
655
|
|
|
656
|
-
Asked by `/multi-agent` and `:local`; autopilot reads `prefs.global.testDepth.default` and degrades to the best open option. Asked here, not Phase
|
|
656
|
+
Asked by `/multi-agent` and `:local`; autopilot reads `prefs.global.testDepth.default` and degrades to the best open option. Asked here, not Phase 3, which four of the eight modes drop.
|
|
657
657
|
|
|
658
658
|
Log: `Phase 0 Step 7.7: testDepth = {unit|unit+ui|unit+mcp} (source {user|autopilot|default|forced}), tier1/tier2 = {open|closed}`
|
|
659
659
|
|
|
@@ -689,7 +689,7 @@ LOG_METRIC_FORWARD_TO_TRACKER=1 $HOME/.claude/scripts/log-metric.sh "$TASK_ID" 0
|
|
|
689
689
|
duration_ms=$D tokens_in=$TI tokens_out=$TO
|
|
690
690
|
```
|
|
691
691
|
|
|
692
|
-
Phase
|
|
692
|
+
Phase 5 cost rollup carries this as a `phase 0` line item so the user sees ambiguity-scoring cost separately from Phase 1 Analysis.
|
|
693
693
|
|
|
694
694
|
<!-- progress-contract: applied -->
|
|
695
695
|
|
|
@@ -710,15 +710,15 @@ node "$HOME/.claude/scripts/usage-register.mjs" --quiet >/dev/null 2>&1 || true
|
|
|
710
710
|
node "$HOME/.claude/scripts/usage-report.mjs" --task-id "$TASK_ID" >/dev/null 2>&1 || true
|
|
711
711
|
```
|
|
712
712
|
|
|
713
|
-
The third line reports the run as started: reporting only from Phase
|
|
714
|
-
only runs that finish, and few do. Phase
|
|
713
|
+
The third line reports the run as started: reporting only from Phase 5 reported
|
|
714
|
+
only runs that finish, and few do. Phase 5 upserts the same key over it. The
|
|
715
715
|
second is the backstop for a machine that reached neither setup nor update - it
|
|
716
716
|
is a no-op once a token resolves, and permanently so under `usageLog.optOut`.
|
|
717
717
|
|
|
718
718
|
It asserts five things, each of which has failed silently in a real run:
|
|
719
719
|
|
|
720
720
|
1. **`agent-state.json` exists.** Every later phase reasons from it.
|
|
721
|
-
2. **`taskType` is set.** Phase
|
|
721
|
+
2. **`taskType` is set.** Phase 2 branches on it (Step 7).
|
|
722
722
|
3. **A Figma reference forces `taskType: "component"`, and `figmaAccess.tier` is
|
|
723
723
|
recorded**, so a later phase can tell a confirmed design from an unfetched one.
|
|
724
724
|
4. **`baseBranchSource` is recorded**, and an interactive run recorded `asked` or
|
|
@@ -730,4 +730,4 @@ Each failed silently in a real run; the script header names which.
|
|
|
730
730
|
|
|
731
731
|
A failure is a halt, not a warning: fix the state, re-run the gate, and leave the phase
|
|
732
732
|
`in_progress` until it passes. Never close Phase 0 on the grounds that its steps ran -
|
|
733
|
-
the gate checks the output, which is what Phase
|
|
733
|
+
the gate checks the output, which is what Phase 2 consumes.
|