@kontextmind/kxm 0.7.95 → 0.7.96

This diff represents the content of publicly available package versions that have been released to one of the supported registries. The information contained in this diff is provided for informational purposes only and reflects changes between package versions as they appear in their respective public registries.
Files changed (140) hide show
  1. package/.claude-plugin/marketplace.json +1 -1
  2. package/.kxm/README.md +39 -9
  3. package/CHANGELOG.md +1 -1
  4. package/README.md +147 -257
  5. package/SECURITY.md +21 -12
  6. package/docs/README.md +133 -54
  7. package/docs/adr/ADR-0002-browser-automation-steel-doks.md +24 -18
  8. package/docs/adr/ADR-0003-sqlite-only-store.md +100 -0
  9. package/docs/adr/ADR-0004-edge-identity-authentik.md +99 -0
  10. package/docs/adr/README.md +33 -0
  11. package/docs/concepts/architecture.md +262 -0
  12. package/docs/concepts/data-and-storage.md +194 -0
  13. package/docs/concepts/trust-model.md +152 -0
  14. package/docs/contracts/README.md +22 -14
  15. package/docs/contracts/effects-and-recovery.md +3 -0
  16. package/docs/contracts/migration.md +2 -2
  17. package/docs/contracts/routing.md +6 -5
  18. package/docs/contributing/assignment-runner.md +388 -0
  19. package/docs/contributing/ci-and-release.md +231 -0
  20. package/docs/contributing/development.md +362 -0
  21. package/docs/contributing/harness-routing-internals.md +192 -0
  22. package/docs/{packages.md → contributing/packages.md} +13 -15
  23. package/docs/{skills → contributing}/repo-work-delivery.md +20 -21
  24. package/docs/contributing/test-matrix.md +208 -0
  25. package/docs/{tui-components.md → contributing/tui-components.md} +30 -22
  26. package/docs/contributing/writing-docs.md +340 -0
  27. package/docs/glossary.md +471 -0
  28. package/docs/guides/agent-skills.md +137 -0
  29. package/docs/guides/browser-automation.md +160 -0
  30. package/docs/guides/context-and-memory.md +352 -0
  31. package/docs/guides/continuous-improvement.md +228 -0
  32. package/docs/guides/governed-skills.md +173 -0
  33. package/docs/guides/nous-providers.md +186 -0
  34. package/docs/guides/peer-messaging.md +304 -0
  35. package/docs/guides/pi-workers.md +219 -0
  36. package/docs/guides/provenance-gates.md +313 -0
  37. package/docs/guides/webhook-workflows.md +364 -0
  38. package/docs/kb/how-credentials-retrieved-safely.md +38 -12
  39. package/docs/kb/how-to-capture-and-annotate-section.md +15 -13
  40. package/docs/kb/how-to-connect-playwright-to-steel.md +16 -11
  41. package/docs/kb/how-to-recover-expired-session-or-orphan.md +26 -16
  42. package/docs/kb/how-to-resume-after-mfa.md +19 -11
  43. package/docs/kb/how-to-take-over-session.md +17 -13
  44. package/docs/kb/why-authentication-disappeared.md +22 -14
  45. package/docs/kb/why-automation-opened-different-browser.md +23 -14
  46. package/docs/kb/why-session-viewer-cannot-control.md +13 -12
  47. package/docs/operations/backup-and-restore.md +248 -0
  48. package/docs/operations/deploy.md +307 -0
  49. package/docs/operations/monitoring.md +209 -0
  50. package/docs/operations/runtime-sync.md +192 -0
  51. package/docs/operations/troubleshooting.md +265 -0
  52. package/docs/operations/upgrade.md +124 -0
  53. package/docs/prompts/browser-annotate-feedback.md +7 -7
  54. package/docs/prompts/browser-diagnose-recover.md +11 -10
  55. package/docs/prompts/browser-explore.md +7 -7
  56. package/docs/prompts/browser-repro-fix.md +7 -7
  57. package/docs/prompts/browser-start.md +12 -11
  58. package/docs/prompts/browser-takeover.md +8 -8
  59. package/docs/{cli-reference.md → reference/cli-reference.md} +83 -41
  60. package/docs/{config-reference.md → reference/config-reference.md} +159 -148
  61. package/docs/reference/configuration.md +299 -0
  62. package/docs/reference/harness-routing.md +508 -0
  63. package/docs/reference/http-api.md +203 -0
  64. package/docs/reference/tools.md +370 -0
  65. package/docs/{workflow-guide.md → reference/workflow-catalog.md} +92 -153
  66. package/docs/reference/workflow-definitions.md +286 -0
  67. package/docs/start/first-workflow.md +287 -0
  68. package/docs/start/install.md +146 -0
  69. package/docs/start/quickstart-claude-code.md +405 -0
  70. package/docs/start/quickstart-pi.md +213 -0
  71. package/docs/templates/README.md +78 -73
  72. package/docs/templates/adr.md +13 -13
  73. package/docs/templates/architecture.md +55 -71
  74. package/docs/templates/bug-fix.md +13 -16
  75. package/docs/templates/feature.md +14 -19
  76. package/docs/templates/handoff.md +44 -46
  77. package/docs/templates/postmortem.md +30 -43
  78. package/docs/templates/research.md +15 -20
  79. package/docs/templates/review.md +49 -50
  80. package/docs/templates/runbook.md +38 -30
  81. package/docs/templates/test-plan.md +16 -23
  82. package/docs/templates/test-report.md +14 -17
  83. package/examples/README.md +9 -5
  84. package/examples/provenance-workflow.json +1 -1
  85. package/examples/webhook-workflows/jira-development.json +59 -0
  86. package/examples/webhook-workflows/jira-issue-updated.json +12 -0
  87. package/package.json +1 -1
  88. package/packages/core/tui/README.md +1 -1
  89. package/plugins/kxm/.claude-plugin/plugin.json +1 -1
  90. package/plugins/kxm/README.md +31 -32
  91. package/plugins/kxm/dist/cli.js +5 -5
  92. package/plugins/kxm/dist/mcp-server.js +1 -1
  93. package/plugins/kxm/dist/runtime.js +1 -1
  94. package/plugins/kxm/package.json +1 -1
  95. package/plugins/kxm/skills/kxm/references/protocol.md +3 -1
  96. package/plugins/kxm/skills/kxm-browser-auth/SKILL.md +1 -1
  97. package/plugins/kxm/skills/kxm-browser-diagnostics/SKILL.md +5 -5
  98. package/plugins/kxm/skills/kxm-browser-explore/SKILL.md +2 -2
  99. package/plugins/kxm/skills/kxm-browser-session/SKILL.md +10 -13
  100. package/plugins/kxm/skills/kxm-browser-takeover/SKILL.md +1 -1
  101. package/plugins/kxm/skills/kxm-browser-verify/SKILL.md +1 -1
  102. package/plugins/kxm/skills/kxm-context-memory/SKILL.md +13 -4
  103. package/plugins/kxm/skills/kxm-hub-ops/SKILL.md +3 -1
  104. package/plugins/kxm/skills/kxm-mind-setup/SKILL.md +2 -1
  105. package/plugins/kxm/skills/kxm-project-setup/SKILL.md +31 -54
  106. package/plugins/kxm/skills/kxm-projects/SKILL.md +1 -1
  107. package/plugins/kxm/skills/kxm-protocol/SKILL.md +1 -1
  108. package/plugins/kxm/skills/kxm-routing-improve/SKILL.md +15 -7
  109. package/plugins/kxm/skills/kxm-runs/SKILL.md +11 -5
  110. package/plugins/kxm/skills/kxm-session/SKILL.md +1 -1
  111. package/plugins/kxm/skills/kxm-tasks/SKILL.md +9 -7
  112. package/plugins/kxm/skills/kxm-workflow/SKILL.md +10 -2
  113. package/plugins/kxm/src/cli/system.ts +1 -1
  114. package/plugins/kxm/src/cli.ts +3 -3
  115. package/plugins/kxm/src/init-guide-setup.ts +1 -1
  116. package/plugins/kxm/src/mcp-server.ts +1 -1
  117. package/plugins/kxm/src/modes.ts +1 -1
  118. package/schemas/README.md +1 -1
  119. package/docs/agent-communication-envelopes-and-gates.md +0 -553
  120. package/docs/agent-skills.md +0 -198
  121. package/docs/architecture.md +0 -245
  122. package/docs/assignment-runner.md +0 -264
  123. package/docs/browser-automation.md +0 -139
  124. package/docs/configuration.md +0 -437
  125. package/docs/continuous-improvement.md +0 -226
  126. package/docs/getting-started.md +0 -277
  127. package/docs/harness-routing.md +0 -616
  128. package/docs/kb/qa-authentik-authentication.md +0 -97
  129. package/docs/kb/qa-extension-install-and-hub-bootstrap.md +0 -85
  130. package/docs/kb/qa-hub-on-a-public-host.md +0 -48
  131. package/docs/kb/qa-sqlite-vs-duckdb.md +0 -35
  132. package/docs/kb/qa-what-the-hub-stores.md +0 -64
  133. package/docs/kxm-handbook.md +0 -1181
  134. package/docs/operations.md +0 -510
  135. package/docs/operator-pi-packages.md +0 -67
  136. package/docs/provenance-gates.md +0 -295
  137. package/docs/skills.md +0 -47
  138. package/docs/test-matrix.md +0 -132
  139. package/docs/troubleshooting.md +0 -322
  140. package/docs/webhook-workflows.md +0 -240
@@ -1,97 +1,49 @@
1
- # Workflow guide
2
-
3
- This Area → Workflow → Stage → Role taxonomy applies to native harness
4
- subscriptions (Claude CLI, Codex, Grok CLI, agy/Antigravity) and to API-key Pi
5
- providers (OpenRouter, Nous Portal). Vendor-prefixed model ids in the candidate
6
- lists are dated research candidates, not admission.
7
-
8
- ## Table of Contents
9
-
10
- 1. [Software Engineering](#software-engineering)
11
- 2. [Design & Experience](#design--experience)
12
- 3. [Media Production](#media-production)
13
- 4. [Data & Analytics](#data--analytics)
14
- 5. [Research & Strategy](#research--strategy)
15
- 6. [Business Operations](#business-operations)
16
- 7. [Security & Reliability](#security--reliability)
17
- 8. [Selection Policy](#selection-policy)
18
- 9. [Slug Registry](#slug-registry)
19
- 10. [Historical Navigation](#historical-navigation)
1
+ # Workflow catalog
20
2
 
21
- ---
22
-
23
- ## Provenance
3
+ This catalog names common multi-agent workflows by area and lists, for each role in them, dated model candidates to evaluate. Use it when you design a workflow: it gives you workflow and role names, a stage breakdown, and a first shortlist of models. It grants nothing; a candidate runs only on a route your project admits, through a harness that is installed and signed in.
24
4
 
25
- The workflow and role candidate lists in this guide are inherited research from the snapshot committed at `67e3f99313dadbbd8ef06884be0588fc92970caf` on 2026-09-06. That is a document snapshot date, not a catalog-verification date; current price and capability freshness is unverified. The lists are dated candidates that require live verification before dispatch. They are not certified prices or capabilities, and they are not an eligibility grant.
5
+ ## Contents
26
6
 
27
- Existing claims inside preserved candidate lines are research claims, not dispatch policy or proven quality guarantees. Area grouping is navigation and never pools unrelated role quality into one global model ranking. Model, harness, platform, modality, required tools, and personal or work context are routing attributes, not area trees.
7
+ 1. [How to read this catalog](#how-to-read-this-catalog)
8
+ 2. [Slug registry](#slug-registry)
9
+ 3. [Documentation for each workflow](#documentation-for-each-workflow)
10
+ 4. [Software Engineering](#software-engineering)
11
+ 5. [Design & Experience](#design--experience)
12
+ 6. [Media Production](#media-production)
13
+ 7. [Data & Analytics](#data--analytics)
14
+ 8. [Research & Strategy](#research--strategy)
15
+ 9. [Business Operations](#business-operations)
16
+ 10. [Security & Reliability](#security--reliability)
28
17
 
29
18
  ---
30
19
 
31
- ## Selection Policy
20
+ ## How to read this catalog
21
+
22
+ The catalog is a four-level taxonomy: **Area → Workflow → Stage → Role**. Each area groups related workflows, each workflow lists its stages in order, and each role belongs to one stage, with a short domain description and up to five ranked candidate models.
32
23
 
33
- Candidate selection is measured per role. Filter stages are optional and are not mandatory Tier-0 gating. Prefer a provider-native authenticated subscription when Tracking says that harness is eligible. For Gemini candidates (`google/*`), the admitted route is the bundled `antigravity` **Pi
34
- provider** (Tracking → Decided, 2026-09-15) — not the OpenRouter provider id, and not a
35
- shell-out to `agy`, which stays a catalog/helper entry. Do not change the candidate ids themselves (they are dated research). Evidence and review remain workflow-specific. Both the Fable architecture critic and the Sol CLI critic remain required for this developer assignment runner.
24
+ - **Candidates are dated research.** The lists come from a documentation snapshot taken on 2026-09-06. Prices (USD per million input and output tokens), context windows and capability notes were not re-verified after that date. Verify them before you dispatch work, and read every quality note as a research claim, not a guarantee.
25
+ - **Candidates are not admission.** A model ID here is a research identifier, not a route selector. Your project admits routes in `.kxm/routes.yaml`, and `kxm harness list` shows which harnesses are installed and signed in. [Harness routing](harness-routing.md) explains when to use a native harness and when an aggregator.
26
+ - **Selection is per role.** Areas are navigation only; they never pool unrelated roles into one ranking. Model, harness, platform, modality and required tools are routing attributes, not areas.
27
+ - **Gates stay deterministic.** Where a role is a critic next to a gate, the gate (tests, type checks, linters) decides pass or fail, and the model reviews what a gate cannot check.
36
28
 
37
- ### Cost band reference (dated candidates)
29
+ An interactive `kxm init` offers to install the Software Engineering workflows when at least one harness is signed in. For each role it picks the first candidate whose harness is installed and signed in, and writes one agent file per role and one `kxm.workflow.v1` file per workflow you choose. Existing files are kept. Set `KXM_SKIP_GUIDE_SETUP_PROMPT=1` to skip the offer.
38
30
 
39
- | Tier | Economic Band | Context Window | Primary Models | Core Workload Profile |
31
+ ### Cost bands (dated candidates)
32
+
33
+ | Tier | Price band | Context window | Example models | Typical work |
40
34
  |---|---|---|---|---|
41
- | **Tier 0: Filter & Ingestion** | $0.03 - $0.20 / M | 1.0M - 1.31M | DeepSeek-V4-Flash, Qwen-3.7-Flash, GLM-5.3-Flash, GPT-5.6-Luna | High-throughput telemetry, raw log ingestion, triage, video pre-filtering, and initial inbox classification. |
42
- | **Tier 1: Workhorse & Engine** | $0.30 - $1.00 / M | 262k - 1.05M | Gemini-3.8-Flash, Qwen-3-Coder-Plus, DeepSeek-V4-Pro, Devstral-2512 | Routine code generation, multimodal visual QA, Text-to-SQL, AST diff patching, and fast tool dispatching. |
43
- | **Tier 2: Precision & Critic** | $1.25 - $4.00 / M | 200k - 1.05M | GPT-5.6-Sol, Grok-4.6, Claude Sonnet 5, OpenAI o3, DeepSeek-R1, Codex 5.3 | Primary code writing, formal contract generation, concurrency race diagnosis, and causal inference. |
44
- | **Tier 3: Sovereign Architecture** | $5.00 - $15.00 / M | 1.0M - 1.05M | Claude Fable 5.1, Claude Opus 5, GPT-6-Astra-Pro (Escalation only) | Designated architecture critic, executive brief, high-stakes legal redlining, and sovereign RFC review. |
45
-
46
- ### Developer runner loop
47
-
48
- The developer assignment runner executes a gated loop with rework across every software engineering workflow. This is not a seventh stage or an informal convention; it is the enforced loop for every software engineering workflow on this runner:
49
-
50
- ```text
51
- plan (claude/fable)
52
- │
53
- ▼
54
- implement (grok-native)
55
- │
56
- ▼
57
- witness (fixed gate: npm run verify)
58
- │
59
- ▼
60
- dual critics (review-arch: fable + review-cli: sol)
61
- │
62
- ├─ If either critic BLOCK ──────────────┐
63
- │ ▼
64
- │ repair (rework_of binding;
65
- │ failover to next eligible
66
- │ authenticated writer)
67
- │ │
68
- │ └─ re-witness + fresh dual review
69
- │
70
- ▼ (both critics PASS on exact tree)
71
- accept (just accept binds commit + both PASS records)
72
- │
73
- ▼
74
- pull request (five CI jobs: 2 Linux validate, classify, docs, plugin)
75
- ```
76
-
77
- **Runner loop invariants:**
78
-
79
- - **Roles and rotation:** The runner strictly maps assignments to authenticated roster roles (`AGENTS.md` and `.kxm/roster.yaml`):
80
- - **Implement / write code (`writer`):** **Grok** (`grok --model grok-4.6`, headless). Currently admitted native writer on this runner. Failover follows attempts-and-relief: after an empty or failed attempt, immediately fail over to the next eligible authenticated writer in the roster (**Qwen** `openrouter/qwen/qwen3-coder-plus` via Pi).
81
- - **Plan (`planner`):** **Claude Fable** (`claude --model fable`), read-only architecture and permissions planning.
82
- - **Review architecture (`reviewer-arch`):** **Claude Fable** (`claude --model fable`), designated architecture and permissions critic.
83
- - **Review CLI / contracts (`reviewer-cli`):** **Codex Sol** (`codex` `gpt-5.6-sol`), designated CLI, protocol, and documentation critic.
84
- - **Fixed witness gate:** `npm run verify` (incorporating tests, typecheck, docs lint, version checks, and generated `dist` match) runs deterministically as the witness gate before critics review. Deterministic gates beat a third model; models propose, gates hold the line. Model critics review contracts, architecture, and documentation; they never replace the deterministic compiler or test runner.
85
- - **Dual-critic quorum:** PR acceptance strictly requires **both** designated critics (**Fable** for architecture/permissions and **Sol** for CLI/contracts) to record a `PASS`. A single critic is at most preliminary triage; neither critic may share a provider with each other or with the writer.
86
- - **Repair back-edge and rework:** If either critic records a `BLOCK`, the runner emits a `repair` assignment bound to the prior attempt via `rework_of`. Every repair round re-runs the fixed witness gate and requires **fresh** dual critic reviews on the resulting tree.
87
- - **Attempts and relief:** Never stop solely because attempts are exhausted or failed. Immediately try the next suggested eligible authenticated model and transfer findings while preserving every attempt, candidate, failed check, and cost record (`AGENTS.md`).
88
- - **Acceptance and auto-merge:** `just accept` binds the exact candidate commit hash and independent Fable + Sol PASS records. Changes are pushed to a feature branch and merged via pull request with auto-merge after all five CI jobs pass.
35
+ | **Tier 0: Filter and ingestion** | $0.03 - $0.20 / M | 1.0M - 1.31M | DeepSeek-V4-Flash, Qwen-3.7-Flash, GLM-5.3-Flash, GPT-5.6-Luna | High-throughput telemetry, raw log ingestion, triage, video pre-filtering, and initial inbox classification. |
36
+ | **Tier 1: Workhorse** | $0.30 - $1.00 / M | 262k - 1.05M | Gemini-3.8-Flash, Qwen-3-Coder-Plus, DeepSeek-V4-Pro, Devstral-2512 | Routine code generation, multimodal visual QA, Text-to-SQL, AST diff patching, and fast tool dispatching. |
37
+ | **Tier 2: Precision and critic** | $1.25 - $4.00 / M | 200k - 1.05M | GPT-5.6-Sol, Grok-4.6, Claude Sonnet 5, OpenAI o3, DeepSeek-R1, Codex 5.3 | Primary code writing, formal contract generation, concurrency race diagnosis, and causal inference. |
38
+ | **Tier 3: Architecture** | $5.00 - $15.00 / M | 1.0M - 1.05M | Claude Fable 5.1, Claude Opus 5, GPT-6-Astra-Pro (escalation only) | Architecture review, executive briefs, high-stakes legal redlining, and RFC review. |
89
39
 
90
40
  ---
91
41
 
92
- ## Slug Registry
42
+ ## Slug registry
43
+
44
+ Area and workflow slugs are lower-case kebab-case ASCII and do not use numerals as canonical identity. They are documentation identity, declared here separately from display names and from GitHub auto-anchors. Role slugs are reusable specialty slugs: kebab-case, no area or workflow prefix, and no numbers.
93
45
 
94
- Area and workflow slugs are lower-case kebab-case ASCII and do not use numerals as canonical identity. They are documentation identity, declared here separately from display names and from GitHub auto-anchors. Role slugs are reusable specialty slugs: kebab-case, no area or workflow prefix, and no numbers. The same specialty may recur across workflows; a role slug must be unique within a workflow. When a role must be disambiguated, the composite reference is the explicit namespace `area-slug/workflow-slug/role-slug` (for example `security-reliability/investigate-incident/forensic-causal-analyst`). Never use a bare number such as 3.1.2 as identity.
46
+ The same specialty may recur across workflows; a role slug must be unique within a workflow. When a role must be disambiguated, the composite reference is the explicit namespace `area-slug/workflow-slug/role-slug` (for example `security-reliability/investigate-incident/forensic-causal-analyst`). Never use a bare number such as 3.1.2 as identity.
95
47
 
96
48
  These slugs imply no runtime config, role admission, schema field, CLI behavior, or alias. Role slugs are declared under each role heading.
97
49
 
@@ -134,6 +86,20 @@ These slugs imply no runtime config, role admission, schema field, CLI behavior,
134
86
  | Security & Reliability | Investigate Incident | `investigate-incident` |
135
87
  | Security & Reliability | Patch Vulnerability | `patch-vulnerability` |
136
88
 
89
+ ## Documentation for each workflow
90
+
91
+ This cross-reference points each software and security workflow at the KXM pages its work most often needs. Like the slugs, it implies no runtime configuration, admission, schema or CLI behavior.
92
+
93
+ | Workflow slug | Primary docs |
94
+ |---|---|
95
+ | `build-feature` | [Set up a first workflow](../start/first-workflow.md), [Architecture](../concepts/architecture.md), [Workflow definitions](workflow-definitions.md), [Test matrix](../contributing/test-matrix.md) |
96
+ | `refactor-repair-regressions` | [Architecture](../concepts/architecture.md), [Test matrix](../contributing/test-matrix.md), [Troubleshooting](../operations/troubleshooting.md), [Provenance gates](../guides/provenance-gates.md) |
97
+ | `stabilize-flaky-tests` | [Test matrix](../contributing/test-matrix.md), [Monitoring](../operations/monitoring.md), [Troubleshooting](../operations/troubleshooting.md) |
98
+ | `design-software-system` | [Architecture](../concepts/architecture.md), [Configuration file reference](config-reference.md), [KXM contracts](../contracts/README.md) |
99
+ | `maintain-documentation` | [Writing docs](../contributing/writing-docs.md), [CLI reference](cli-reference.md) |
100
+ | `investigate-incident` | [Monitoring](../operations/monitoring.md), [Troubleshooting](../operations/troubleshooting.md), [Webhook workflows](../guides/webhook-workflows.md) |
101
+ | `patch-vulnerability` | [Provenance gates](../guides/provenance-gates.md), [Trust model](../concepts/trust-model.md), [Environment variables and limits](configuration.md) |
102
+
137
103
  ---
138
104
 
139
105
  ## Software Engineering
@@ -142,11 +108,11 @@ These slugs imply no runtime config, role admission, schema field, CLI behavior,
142
108
 
143
109
  **Overview:**
144
110
 
145
- 1. **Build Feature:** PRD Ingestion $\to$ System Planning $\to$ API Spec/Contract Authoring $\to$ Workspace Scaffolding $\to$ Core Backend & Data Logic $\to$ Full-Stack UI Implementation $\to$ Tool/SDK Integrations $\to$ Automated Verification.
146
- 2. **Refactor and Repair Regressions:** Codebase Smell Analysis $\to$ Modular Decomposition $\to$ Surgical Multi-File AST Transforms $\to$ Static Typing & Contract Verification.
147
- 3. **Stabilize Flaky Tests:** Flakiness Root-Cause Extraction $\to$ Deterministic Mock/Async Hardening.
148
- 4. **Design Software System:** NFR/SLA Ingestion $\to$ System Topology & Trade-Offs $\to$ Technical RFC Authoring $\to$ STRIDE Threat Modeling $\to$ Storage/Sharding Design $\to$ IaC Cloud Topology $\to$ Chaos/DR Review $\to$ Independent Critic Quorum.
149
- 5. **Maintain Documentation:** Accuracy Audit $\to$ Documentation Write $\to$ Witness Gate $\to$ Dual Review Quorum.
111
+ 1. **Build Feature:** PRD Ingestion → System Planning → API Spec/Contract Authoring → Workspace Scaffolding → Core Backend & Data Logic → Full-Stack UI Implementation → Tool/SDK Integrations → Automated Verification.
112
+ 2. **Refactor and Repair Regressions:** Codebase Smell Analysis → Modular Decomposition → Surgical Multi-File AST Transforms → Static Typing & Contract Verification.
113
+ 3. **Stabilize Flaky Tests:** Flakiness Root-Cause Extraction → Deterministic Mock/Async Hardening.
114
+ 4. **Design Software System:** NFR/SLA Ingestion → System Topology & Trade-Offs → Technical RFC Authoring → STRIDE Threat Modeling → Storage/Sharding Design → IaC Cloud Topology → Chaos/DR Review → Independent Critic Quorum.
115
+ 5. **Maintain Documentation:** Accuracy Audit → Documentation Write → Review.
150
116
 
151
117
  ---
152
118
 
@@ -162,7 +128,7 @@ These slugs imply no runtime config, role admission, schema field, CLI behavior,
162
128
 
163
129
  *Domain:* Decomposes requirements into an executable, dependency-ordered DAG of modules, tasks, and verification gates.
164
130
 
165
- 1. **`anthropic/claude-fable-5.1`** (1M ctx | $10.00 / $50.00) — Sovereign systems planner; strict boundary isolation and non-overlapping DAG synthesis.
131
+ 1. **`anthropic/claude-fable-5.1`** (1M ctx | $10.00 / $50.00) — Systems planning; strict boundary isolation and non-overlapping DAG synthesis.
166
132
  2. **`deepseek/deepseek-v4-pro-0813`** (1.05M ctx | $0.58 / $1.74) — Exceptional price-to-reasoning ratio; detects race conditions and circular dependencies.
167
133
  3. **`openai/gpt-6-astra-pro`** (1.05M ctx | $10.00 / $50.00) — Enterprise task decomposition and critical-path scheduling across massive PRDs.
168
134
  4. **`x-ai/grok-4.6`** (500k ctx | $2.00 / $6.00) — Pragmatic engineering focus; rapid modular task decomposition.
@@ -186,7 +152,7 @@ These slugs imply no runtime config, role admission, schema field, CLI behavior,
186
152
 
187
153
  *Domain:* Provisions monorepo workspace topologies (Turborepo, Cargo, pnpm), linters, Docker multi-stage builds, and CI pipelines.
188
154
 
189
- 1. **`x-ai/grok-4.6`** (500k ctx | $2.00 / $6.00) — Designated native repo builder; resolves path aliases, linkings, and workspace configs cleanly.
155
+ 1. **`x-ai/grok-4.6`** (500k ctx | $2.00 / $6.00) — Fast repository builder; resolves path aliases, links, and workspace configs cleanly.
190
156
  2. **`mistralai/devstral-2512`** (262k ctx | $0.40 / $2.00) — Developer-focused open model; produces clean Makefiles, Taskfiles, and tool configs.
191
157
  3. **`qwen/qwen3-coder-plus`** (1M ctx | $0.65 / $3.25) — Coordinates cross-package build configs and dependency lockfiles in 1M context.
192
158
  4. **`openai/gpt-5.3-codex`** (400k ctx | $1.75 / $14.00) — Complex CI/CD workflows and multi-stage container build optimizations.
@@ -198,7 +164,7 @@ These slugs imply no runtime config, role admission, schema field, CLI behavior,
198
164
 
199
165
  *Domain:* Implements domain entities, SQL queries/migrations, ORM persistence, concurrency primitives, and transactional boundaries.
200
166
 
201
- 1. **`x-ai/grok-4.6`** (500k ctx | $2.00 / $6.00) — Admitted native writer; passes deterministic verify gates (`npm run verify`) on the first pass with minimal rework.
167
+ 1. **`x-ai/grok-4.6`** (500k ctx | $2.00 / $6.00) — Strong first-pass implementation that passes deterministic verification gates with minimal rework.
202
168
  2. **`qwen/qwen3-coder-plus`** (1M ctx | $0.65 / $3.25) — 1M-context open-weight anchor; generates performant database layers and CRUD services.
203
169
  3. **`deepseek/deepseek-v4-pro-0813`** (1.05M ctx | $0.58 / $1.74) — Exceptional database logic, indexing, and edge-case domain implementation.
204
170
  4. **`openai/gpt-5.3-codex`** (400k ctx | $1.75 / $14.00) — Algorithmic precision in Rust, Go, Python, and TypeScript concurrent systems.
@@ -254,7 +220,7 @@ These slugs imply no runtime config, role admission, schema field, CLI behavior,
254
220
 
255
221
  *Domain:* Executes surgical multi-file edits, applying design patterns while preserving comments and formatting.
256
222
 
257
- 1. **`x-ai/grok-4.6`** (500k ctx | $2.00 / $6.00) — Designated native writer; fast execution of multi-file semantic code changes.
223
+ 1. **`x-ai/grok-4.6`** (500k ctx | $2.00 / $6.00) — Fast execution of multi-file semantic code changes.
258
224
  2. **`relace/relace-apply-3`** (256k ctx | $0.85 / $1.25) — Specialized model for deterministic AST code patch merging without syntax corruption.
259
225
  3. **`openai/gpt-5.3-codex`** (400k ctx | $1.75 / $14.00) — High-precision refactoring of complex pointer/generic logic.
260
226
  4. **`morph/morph-v3-large`** (262k ctx | $0.90 / $1.90) — 4,500 tok/sec high-accuracy mechanical patch application.
@@ -264,9 +230,9 @@ These slugs imply no runtime config, role admission, schema field, CLI behavior,
264
230
 
265
231
  *Slug:* `semantic-equivalence-verifier` | *Stage:* Equivalence verification
266
232
 
267
- *Domain:* Critic-with-gate. Deterministic checks (`tsc --noEmit`, `npm test`, `npm run check`, generated `dist` match) serve as the fixed witness gate and are never delegated to an LLM; the model role reviews public contract compliance, CLI behavior, and behavioral equivalence without replacing the gate ("Deterministic gates beat a third model… models propose; gates hold the line").
233
+ *Domain:* Critic-with-gate: reviews public contract compliance and behavioral equivalence; deterministic checks (type checks, tests, build output) remain the fixed gate and are never delegated to a model.
268
234
 
269
- 1. **`openai/gpt-5.6-sol`** (1.05M ctx | $2.00 / $10.00) — Designated CLI/spec critic; reviews compiler cascades and type-checker cascades against API contracts.
235
+ 1. **`openai/gpt-5.6-sol`** (1.05M ctx | $2.00 / $10.00) — Reviews compiler and type-checker cascades against API contracts.
270
236
  2. **`openai/o3-mini-high`** (200k ctx | $1.10 / $4.40) — Deep symbolic verification that AST refactors maintain behavioral equivalence.
271
237
  3. **`deepseek/deepseek-v4-pro-0813`** (1.05M ctx | $0.58 / $1.74) — Type-cascade and semantic regression review across package boundaries.
272
238
  4. **`anthropic/claude-opus-5`** (1M ctx | $5.00 / $25.00) — Verifies backwards compatibility and deprecation notices across public APIs.
@@ -306,7 +272,7 @@ These slugs imply no runtime config, role admission, schema field, CLI behavior,
306
272
 
307
273
  *Domain:* Evaluates CAP/PACELC trade-offs, consensus protocols (Raft, Sagas), partition boundaries, and authors primary RFCs.
308
274
 
309
- 1. **`anthropic/claude-fable-5.1`** (1M ctx | $10.00 / $50.00) — Designated architecture planner; strict fault domain and consensus design.
275
+ 1. **`anthropic/claude-fable-5.1`** (1M ctx | $10.00 / $50.00) — Strict fault-domain and consensus design.
310
276
  2. **`openai/o3`** (200k ctx | $2.00 / $8.00) — Formal proof of consensus safety, liveness, and split-brain recovery logic.
311
277
  3. **`openai/gpt-6-astra-pro`** (1.05M ctx | $10.00 / $50.00) — Multi-tier cloud topology and global multi-region architecture design.
312
278
  4. **`anthropic/claude-opus-5`** (1M ctx | $5.00 / $25.00) — Publication-grade RFC authoring detailing operational trade-offs and team boundaries.
@@ -330,8 +296,8 @@ These slugs imply no runtime config, role admission, schema field, CLI behavior,
330
296
 
331
297
  *Domain:* Conducts STRIDE threat assessments, reviews IAM least-privilege policies, mTLS boundaries, and zero-trust architectures.
332
298
 
333
- 1. **`anthropic/claude-fable-5.1`** (1M ctx | $10.00 / $50.00) — Designated architecture/permissions critic; zero-trust network modeling.
334
- 2. **`openai/gpt-5.6-sol`** (1.05M ctx | $2.00 / $10.00) — Designated CLI/spec critic; auth token and protocol threat analysis.
299
+ 1. **`anthropic/claude-fable-5.1`** (1M ctx | $10.00 / $50.00) — Architecture and permissions review; zero-trust network modeling.
300
+ 2. **`openai/gpt-5.6-sol`** (1.05M ctx | $2.00 / $10.00) — Auth token and protocol threat analysis.
335
301
  3. **`openai/gpt-6-astra-pro`** (1.05M ctx | $10.00 / $50.00) — Enterprise IAM and cloud privilege escalation modeling.
336
302
  4. **`anthropic/claude-opus-5`** (1M ctx | $5.00 / $25.00) — STRIDE threat matrix synthesis and compliance verification.
337
303
  5. **`deepseek/deepseek-v4-pro-0813`** (1.05M ctx | $0.58 / $1.74) — Full-infrastructure vulnerability surface analysis.
@@ -366,8 +332,8 @@ These slugs imply no runtime config, role admission, schema field, CLI behavior,
366
332
 
367
333
  *Domain:* Independent peer-review; challenges unvalidated assumptions, prevents architectural drift, and verifies contracts.
368
334
 
369
- 1. **`anthropic/claude-fable-5.1`** (1M ctx | $10.00 / $50.00) — **Designated Architecture & Permissions Critic** (Mandatory for PR acceptance).
370
- 2. **`openai/gpt-5.6-sol`** (1.05M ctx | $2.00 / $10.00) — **Designated CLI, Protocol & Contracts Critic** (Mandatory for PR acceptance).
335
+ 1. **`anthropic/claude-fable-5.1`** (1M ctx | $10.00 / $50.00) — Architecture and permissions criticism.
336
+ 2. **`openai/gpt-5.6-sol`** (1.05M ctx | $2.00 / $10.00) — CLI, protocol and contract criticism.
371
337
  3. **`z-ai/glm-5.3`** (1.31M ctx | $1.40 / $4.40) — 1.31M context full-stack protocol audit.
372
338
  4. **`openai/gpt-6-astra-pro`** (1.05M ctx | $10.00 / $50.00) — Enterprise RFC compliance verification.
373
339
  5. **`qwen/qwen3.8-max-0902`** (1M ctx | $2.00 / $6.00) — Independent open-weights architectural validation.
@@ -384,9 +350,9 @@ These slugs imply no runtime config, role admission, schema field, CLI behavior,
384
350
 
385
351
  *Slug:* `documentation-accuracy-auditor` | *Stage:* Accuracy audit
386
352
 
387
- *Domain:* Read-only planner; audits developer and operator documentation against CLI `--help` outputs, source environment variables, tool schemas, and repository configuration before documentation authoring begins.
353
+ *Domain:* Read-only audit of developer and operator documentation against CLI `--help` output, source environment variables, tool schemas, and repository configuration before writing begins.
388
354
 
389
- 1. **`anthropic/claude-fable-5.1`** (1M ctx | $10.00 / $50.00) — Designated architecture and permissions planner; audits authority, permissions, and lifecycle models against reality.
355
+ 1. **`anthropic/claude-fable-5.1`** (1M ctx | $10.00 / $50.00) — Audits authority, permissions, and lifecycle models against actual behavior.
390
356
  2. **`deepseek/deepseek-v4-pro-0813`** (1.05M ctx | $0.58 / $1.74) — High-throughput cross-referencing between source code symbols and markdown documentation.
391
357
  3. **`openai/gpt-5.6-sol`** (1.05M ctx | $2.00 / $10.00) — Precision auditing of CLI commands, flags, schema definitions, and environment variables.
392
358
  4. **`qwen/qwen3.8-max-0902`** (1M ctx | $2.00 / $6.00) — Broad codebase auditing against guides, tutorials, and operational manuals.
@@ -398,7 +364,7 @@ These slugs imply no runtime config, role admission, schema field, CLI behavior,
398
364
 
399
365
  *Domain:* Authors concise, accurate developer and operator documentation, ensuring clear terminology, correct command flags, and consistent formatting without marketing drift.
400
366
 
401
- 1. **`x-ai/grok-4.6`** (500k ctx | $2.00 / $6.00) — Admitted native writer; rapid, structured markdown authoring adhering to strict repository conventions.
367
+ 1. **`x-ai/grok-4.6`** (500k ctx | $2.00 / $6.00) — Rapid, structured Markdown authoring that follows repository conventions.
402
368
  2. **`qwen/qwen3-coder-plus`** (1M ctx | $0.65 / $3.25) — High-context technical writer; coordinates cross-file references across large documentation sets.
403
369
  3. **`anthropic/claude-sonnet-5`** (1M ctx | $2.00 / $10.00) — Fluid technical documentation authoring with clear instructional hierarchy.
404
370
  4. **`openai/gpt-5.3-codex`** (400k ctx | $1.75 / $14.00) — Accurate technical guides, CLI flag references, and runnable examples.
@@ -408,10 +374,10 @@ These slugs imply no runtime config, role admission, schema field, CLI behavior,
408
374
 
409
375
  *Slug:* `documentation-reviewer` | *Stage:* Review
410
376
 
411
- *Domain:* Critic-with-gate. Deterministic gates (`npm run lint:docs`, `test/docs-copy.test.ts` docs brake, and `npm run check`) serve as the fixed witness gate and are never replaced by LLMs; critics conduct independent peer review where the architecture critic verifies authority and permission wording, and the CLI critic verifies command/flag accuracy and documentation linting.
377
+ *Domain:* Critic-with-gate: the architecture critic verifies authority and permission wording; the CLI critic verifies command and flag accuracy; a deterministic docs lint remains the fixed gate.
412
378
 
413
- 1. **`anthropic/claude-fable-5.1`** (1M ctx | $10.00 / $50.00) — **Designated Architecture & Permissions Critic** (Mandatory for PR acceptance); verifies security, governance, and authority wording.
414
- 2. **`openai/gpt-5.6-sol`** (1.05M ctx | $2.00 / $10.00) — **Designated CLI, Protocol & Contracts Critic** (Mandatory for PR acceptance); verifies CLI command accuracy, flag syntax, and documentation formatting.
379
+ 1. **`anthropic/claude-fable-5.1`** (1M ctx | $10.00 / $50.00) — Verifies security, governance, and authority wording.
380
+ 2. **`openai/gpt-5.6-sol`** (1.05M ctx | $2.00 / $10.00) — Verifies CLI command accuracy, flag syntax, and documentation formatting.
415
381
  3. **`anthropic/claude-opus-5`** (1M ctx | $5.00 / $25.00) — Rigorous editorial review for clarity, technical completeness, and tone consistency.
416
382
  4. **`deepseek/deepseek-v4-pro-0813`** (1.05M ctx | $0.58 / $1.74) — Systematic verification of cross-document links and anchor consistency.
417
383
  5. **`mistralai/devstral-2512`** (262k ctx | $0.40 / $2.00) — Rapid lint rule compliance and syntax verification.
@@ -424,10 +390,10 @@ These slugs imply no runtime config, role admission, schema field, CLI behavior,
424
390
 
425
391
  **Overview:**
426
392
 
427
- 1. **Build Design System:** W3C DTCG Token Modeling $\to$ Style Dictionary v4 AST Engine $\to$ Headless Component Primitives $\to$ Compound Variants (CVA).
428
- 2. **Engineer Mobile Interactions:** Harmonic Spring Physics $\to$ Reanimated v3 UI-Thread Worklets $\to$ SwiftUI/Compose Gestures $\to$ CoreHaptics Waveforms $\to$ 120Hz LTPO Profiling.
429
- 3. **Engineer Terminal Interfaces:** Raw Mode Protocol $\to$ TrueColor Double-Buffering $\to$ The Elm Architecture (Bubbletea) / Ratatui Layouts $\to$ Terminal Signal Trapping.
430
- 4. **Audit Visual Quality and Accessibility:** Snapshot Ingestion $\to$ Perceptual Diffing (SSIM) $\to$ WCAG 2.2 / APCA Contrast Math $\to$ Concentric Radius / Optical Polish.
393
+ 1. **Build Design System:** W3C DTCG Token Modeling → Style Dictionary v4 AST Engine → Headless Component Primitives → Compound Variants (CVA).
394
+ 2. **Engineer Mobile Interactions:** Harmonic Spring Physics → Reanimated v3 UI-Thread Worklets → SwiftUI/Compose Gestures → CoreHaptics Waveforms → 120Hz LTPO Profiling.
395
+ 3. **Engineer Terminal Interfaces:** Raw Mode Protocol → TrueColor Double-Buffering → The Elm Architecture (Bubbletea) / Ratatui Layouts → Terminal Signal Trapping.
396
+ 4. **Audit Visual Quality and Accessibility:** Snapshot Ingestion → Perceptual Diffing (SSIM) → WCAG 2.2 / APCA Contrast Math → Concentric Radius / Optical Polish.
431
397
 
432
398
  ---
433
399
 
@@ -509,7 +475,7 @@ These slugs imply no runtime config, role admission, schema field, CLI behavior,
509
475
 
510
476
  1. **`x-ai/grok-4.6`** (500k ctx | $2.00 / $6.00) — Pragmatic terminal systems code writer; adheres to strict TUI layout constraints.
511
477
  2. **`openai/gpt-5.3-codex`** (400k ctx | $1.75 / $14.00) — Complex ANSI rendering engines and terminal buffer diffing.
512
- 3. **`poolside/laguna-s-2.1`** (1.05M ctx | $0.09 / $0.18) — **Terminal-Bench 2.1 Specialist (70.2% score)**; 1M context at $0.09/M.
478
+ 3. **`poolside/laguna-s-2.1`** (1.05M ctx | $0.09 / $0.18) — Low-cost candidate; 1M context at $0.09/M.
513
479
  4. **`mistralai/devstral-2512`** (262k ctx | $0.40 / $2.00) — Clean Go (Bubbletea) and Rust (Ratatui) view/update implementations.
514
480
  5. **`anthropic/claude-opus-5`** (1M ctx | $5.00 / $25.00) — High-ergonomic terminal UX and keyboard navigation design.
515
481
 
@@ -565,9 +531,9 @@ These slugs imply no runtime config, role admission, schema field, CLI behavior,
565
531
 
566
532
  **Overview:**
567
533
 
568
- 1. **Produce Generative Video:** Ideation $\to$ Scriptwriting $\to$ Shot Storyboarding $\to$ Prompt Synthesis $\to$ Video Generation $\to$ Visual Continuity QA $\to$ Audio/Foley Scoring $\to$ Final Render.
569
- 2. **Repurpose Long-Form Video:** Raw Media Ingestion (1–4h) $\to$ Semantic Hook Extraction $\to$ Active Speaker Tracking & Dynamic 9:16 Reframing $\to$ B-Roll Matching $\to$ Kinetic Typography & Packaging.
570
- 3. **Edit and Render Media:** Codec/Asset Ingestion $\to$ Rough-Cut Compilation $\to$ Deterministic NLE/FFmpeg Filtergraph Assembly $\to$ Color LUT/Loudness Normalization $\to$ Broadcast Compliance Gate.
534
+ 1. **Produce Generative Video:** Ideation → Scriptwriting → Shot Storyboarding → Prompt Synthesis → Video Generation → Visual Continuity QA → Audio/Foley Scoring → Final Render.
535
+ 2. **Repurpose Long-Form Video:** Raw Media Ingestion (1–4h) → Semantic Hook Extraction → Active Speaker Tracking & Dynamic 9:16 Reframing → B-Roll Matching → Kinetic Typography & Packaging.
536
+ 3. **Edit and Render Media:** Codec/Asset Ingestion → Rough-Cut Compilation → Deterministic NLE/FFmpeg Filtergraph Assembly → Color LUT/Loudness Normalization → Broadcast Compliance Gate.
571
537
 
572
538
  ---
573
539
 
@@ -721,9 +687,9 @@ These slugs imply no runtime config, role admission, schema field, CLI behavior,
721
687
 
722
688
  **Overview:**
723
689
 
724
- 1. **Query Business Intelligence:** Metric Resolution $\to$ SQL Generation $\to$ AST Optimization $\to$ Warehouse Execution $\to$ Declarative Dashboard Formatting.
725
- 2. **Analyze Dataset:** Data Hygiene $\to$ Hypothesis Formulation $\to$ Vectorized Python (Polars/Pandas) $\to$ Sandbox Execution $\to$ Causal Inference.
726
- 3. **Extract and Audit Documents:** Massive 10-K Ingestion $\to$ Table/Chart Extraction $\to$ High-Precision Mathematical Auditing $\to$ Executive Variance Memo.
690
+ 1. **Query Business Intelligence:** Metric Resolution → SQL Generation → AST Optimization → Warehouse Execution → Declarative Dashboard Formatting.
691
+ 2. **Analyze Dataset:** Data Hygiene → Hypothesis Formulation → Vectorized Python (Polars/Pandas) → Sandbox Execution → Causal Inference.
692
+ 3. **Extract and Audit Documents:** Massive 10-K Ingestion → Table/Chart Extraction → High-Precision Mathematical Auditing → Executive Variance Memo.
727
693
 
728
694
  ---
729
695
 
@@ -841,8 +807,8 @@ These slugs imply no runtime config, role admission, schema field, CLI behavior,
841
807
 
842
808
  **Overview:**
843
809
 
844
- 1. **Prepare Decision Brief:** Status Ingestion $\to$ Signal Extraction $\to$ Scenario Trade-Off Modeling $\to$ C-Suite Ghostwriting $\to$ Tone Calibration.
845
- 2. **Structure Negotiations:** Contract/Ticket Ingestion $\to$ Leverage (BATNA) Modeling $\to$ Counter-Proposal Structuring.
810
+ 1. **Prepare Decision Brief:** Status Ingestion → Signal Extraction → Scenario Trade-Off Modeling → C-Suite Ghostwriting → Tone Calibration.
811
+ 2. **Structure Negotiations:** Contract/Ticket Ingestion → Leverage (BATNA) Modeling → Counter-Proposal Structuring.
846
812
 
847
813
  ---
848
814
 
@@ -916,9 +882,9 @@ These slugs imply no runtime config, role admission, schema field, CLI behavior,
916
882
 
917
883
  **Overview:**
918
884
 
919
- 1. **Automate Tasks:** Multi-Channel Ingestion $\to$ Urgency Triage $\to$ Action Item Extraction $\to$ Multi-Tool API Dispatch $\to$ Daily Briefing.
920
- 2. **Handle Customer Escalations:** Ticket Ingestion $\to$ De-escalation Comms.
921
- 3. **Review Contracts:** Contract Ingestion $\to$ SLA/Legal Risk Audit $\to$ Redline.
885
+ 1. **Automate Tasks:** Multi-Channel Ingestion → Urgency Triage → Action Item Extraction → Multi-Tool API Dispatch → Daily Briefing.
886
+ 2. **Handle Customer Escalations:** Ticket Ingestion → De-escalation Comms.
887
+ 3. **Review Contracts:** Contract Ingestion → SLA/Legal Risk Audit → Redline.
922
888
 
923
889
  ---
924
890
 
@@ -1012,8 +978,8 @@ These slugs imply no runtime config, role admission, schema field, CLI behavior,
1012
978
 
1013
979
  **Overview:**
1014
980
 
1015
- 1. **Investigate Incident:** Telemetry filtering $\to$ Incident triage $\to$ Forensic analysis $\to$ Reproduction $\to$ Debrief. Stages are optional per incident; there is no mandatory five-model fanout.
1016
- 2. **Patch Vulnerability:** Taint Flow Analysis $\to$ Defensive Patching $\to$ Adversarial Mutation Verification.
981
+ 1. **Investigate Incident:** Telemetry filtering → Incident triage → Forensic analysis → Reproduction → Debrief. Stages are optional per incident; there is no mandatory five-model fanout.
982
+ 2. **Patch Vulnerability:** Taint Flow Analysis → Defensive Patching → Adversarial Mutation Verification.
1017
983
 
1018
984
  ---
1019
985
 
@@ -1098,7 +1064,7 @@ These slugs imply no runtime config, role admission, schema field, CLI behavior,
1098
1064
  *Domain:* Synthesizes fail-closed security fixes (parameterized queries, sanitization, constant-time comparisons).
1099
1065
 
1100
1066
  1. **`openai/gpt-5.3-codex`** (400k ctx | $1.75 / $14.00) — Surgical security patch generator using secure coding primitives.
1101
- 2. **`x-ai/grok-4.6`** (500k ctx | $2.00 / $6.00) — Rapid patch application that passes `npm run verify` cleanly.
1067
+ 2. **`x-ai/grok-4.6`** (500k ctx | $2.00 / $6.00) — Rapid patch application that passes deterministic verification cleanly.
1102
1068
  3. **`qwen/qwen3-coder-plus`** (1M ctx | $0.65 / $3.25) — High-throughput defensive refactoring across large repos.
1103
1069
  4. **`anthropic/claude-sonnet-5`** (1M ctx | $2.00 / $10.00) — Applies defensive validation without breaking valid user flows.
1104
1070
  5. **`bytedance-seed/seed-2.0-code`** (262k ctx | $0.50 / $3.00) — Web and client-side sanitization hardening.
@@ -1117,36 +1083,9 @@ These slugs imply no runtime config, role admission, schema field, CLI behavior,
1117
1083
 
1118
1084
  ---
1119
1085
 
1120
- ## Historical Navigation
1086
+ ## Related
1121
1087
 
1122
- This table is historical navigation from the 2026-09-06 heading numbers to the current documentation slugs. It is not a runtime alias lane.
1123
-
1124
- | Old heading | New workflow slug | Old roles |
1125
- |---|---|---|
1126
- | Workflow 1.1: Generative AI Video Production Pipeline | `produce-generative-video` | 1.1.1–1.1.5 |
1127
- | Workflow 1.2: Automated Long-to-Shorts Content Repurposing Pipeline | `repurpose-long-form-video` | 1.2.1–1.2.3 |
1128
- | Workflow 1.3: Programmatic Post-Production & Tool-Calling Assembly Pipeline | `edit-render-media` | 1.3.1–1.3.2 |
1129
- | Workflow 2.1: Greenfield Software Engineering & Monorepo Scaffolding | `build-feature` | 2.1.1–2.1.6 |
1130
- | Workflow 3.1: Production Incident RCA & Crash Diagnostics | `investigate-incident` | 3.1.1–3.1.3 |
1131
- | Workflow 3.2: Precision Code Refactoring & Regression Repair | `refactor-repair-regressions` | 3.2.1–3.2.3 |
1132
- | Workflow 3.3: Flaky Test Remediation & Vulnerability Patching | (split; see role rows) | 3.3.1–3.3.3 |
1133
- | Role 3.3.1: Concurrency & Flakiness Detective | `stabilize-flaky-tests` | 3.3.1 |
1134
- | Role 3.3.2: Defensive Patch & Hardening Engineer | `patch-vulnerability` | 3.3.2 |
1135
- | Role 3.3.3: Adversarial Security & Mutation Auditor | `patch-vulnerability` | 3.3.3 |
1136
- | Workflow 4.1: Executive Decision Support & Strategic Planning | `prepare-decision-brief` | 4.1.1–4.1.3 |
1137
- | Workflow 4.2: Autonomous Personal Productivity & Task Automation | `automate-tasks` | 4.2.1–4.2.3 |
1138
- | Workflow 4.3: High-Stakes Customer Escalation, Negotiation & Contracts | (split; see role rows) | 4.3.1–4.3.3 |
1139
- | Role 4.3.1: Contract Redline & Commercial Terms Auditor | `review-contracts` | 4.3.1 |
1140
- | Role 4.3.2: Crisis & Customer De-escalation Communicator | `handle-customer-escalations` | 4.3.2 |
1141
- | Role 4.3.3: Negotiation Leverage & Deal Structuring Strategist | `structure-negotiations` | 4.3.3 |
1142
- | Workflow 5.1: Distributed Systems Design & Technical RFC Formulation | `design-software-system` | 5.1.1–5.1.6 |
1143
- | Workflow 6.1: Enterprise Business Intelligence & Text-to-SQL Pipeline | `query-business-intelligence` | 6.1.1–6.1.3 |
1144
- | Workflow 6.2: Automated Data Science & Statistical Modeling | `analyze-dataset` | 6.2.1–6.2.2 |
1145
- | Workflow 6.3: Multimodal Financial & Operational Document Intelligence | `extract-audit-documents` | 6.3.1–6.3.2 |
1146
- | Workflow 6.4: Real-Time Operational Telemetry & Distributed Systems RCA | `investigate-incident` | 6.4.1–6.4.2 |
1147
- | Workflow 7.1: Design System Architecture & Multi-Platform Component Engineering | `build-design-system` | 7.1.1–7.1.2 |
1148
- | Workflow 7.2: Mobile-First Interactive UX & Gesture/Haptic Engineering | `engineer-mobile-interactions` | 7.2.1–7.2.2 |
1149
- | Workflow 7.3: Terminal User Interface (TUI) & Rich CLI Experience Engineering | `engineer-terminal-interfaces` | 7.3.1–7.3.2 |
1150
- | Workflow 7.4: Multimodal Visual Design QA, Accessibility (a11y) & Polish Audit | `audit-visual-accessibility` | 7.4.1–7.4.2 |
1151
-
1152
- Role rotation, admission, and dispatch policy live in AGENTS.md and Tracking, not in this guide.
1088
+ - [Workflow definitions](workflow-definitions.md): write a webhook or Runtime workflow for these stages
1089
+ - [Harness routing](harness-routing.md): choose the route that runs a candidate model
1090
+ - [Configuration file reference](config-reference.md): agents, roles, routes and prices
1091
+ - [Set up a first workflow](../start/first-workflow.md): run a workflow end to end