codex-orchestrator 2.0.11 → 2.0.12

This diff represents the content of publicly available package versions that have been released to one of the supported registries. The information contained in this diff is provided for informational purposes only and reflects changes between package versions as they appear in their respective public registries.
Files changed (262) hide show
  1. package/CHANGELOG.md +15 -0
  2. package/README.md +25 -51
  3. package/dist/src/index.d.ts +2 -8
  4. package/dist/src/index.d.ts.map +1 -1
  5. package/dist/src/index.js +1 -4
  6. package/dist/src/index.js.map +1 -1
  7. package/dist/src/v2/acceptance-proof.d.ts +46 -31
  8. package/dist/src/v2/acceptance-proof.d.ts.map +1 -1
  9. package/dist/src/v2/acceptance-proof.js +157 -195
  10. package/dist/src/v2/acceptance-proof.js.map +1 -1
  11. package/dist/src/v2/active-attempt.d.ts +94 -0
  12. package/dist/src/v2/active-attempt.d.ts.map +1 -0
  13. package/dist/src/v2/active-attempt.js +200 -0
  14. package/dist/src/v2/active-attempt.js.map +1 -0
  15. package/dist/src/v2/adapters/command.d.ts +6 -0
  16. package/dist/src/v2/adapters/command.d.ts.map +1 -1
  17. package/dist/src/v2/adapters/command.js +43 -2
  18. package/dist/src/v2/adapters/command.js.map +1 -1
  19. package/dist/src/v2/candidate.d.ts +15 -31
  20. package/dist/src/v2/candidate.d.ts.map +1 -1
  21. package/dist/src/v2/candidate.js +7 -29
  22. package/dist/src/v2/candidate.js.map +1 -1
  23. package/dist/src/v2/checked-change.d.ts +3 -2
  24. package/dist/src/v2/checked-change.d.ts.map +1 -1
  25. package/dist/src/v2/checked-change.js +4 -3
  26. package/dist/src/v2/checked-change.js.map +1 -1
  27. package/dist/src/v2/cli-contract.d.ts +1 -1
  28. package/dist/src/v2/cli-contract.d.ts.map +1 -1
  29. package/dist/src/v2/cli-contract.js +4 -6
  30. package/dist/src/v2/cli-contract.js.map +1 -1
  31. package/dist/src/v2/cli.d.ts +8 -0
  32. package/dist/src/v2/cli.d.ts.map +1 -1
  33. package/dist/src/v2/cli.js +13 -0
  34. package/dist/src/v2/cli.js.map +1 -1
  35. package/dist/src/v2/code-review-report.d.ts +10 -18
  36. package/dist/src/v2/code-review-report.d.ts.map +1 -1
  37. package/dist/src/v2/code-review-report.js +63 -60
  38. package/dist/src/v2/code-review-report.js.map +1 -1
  39. package/dist/src/v2/codex-process.d.ts +6 -2
  40. package/dist/src/v2/codex-process.d.ts.map +1 -1
  41. package/dist/src/v2/codex-process.js +25 -9
  42. package/dist/src/v2/codex-process.js.map +1 -1
  43. package/dist/src/v2/config.d.ts +0 -2
  44. package/dist/src/v2/config.d.ts.map +1 -1
  45. package/dist/src/v2/config.js +3 -6
  46. package/dist/src/v2/config.js.map +1 -1
  47. package/dist/src/v2/contained-report-operation.d.ts +41 -196
  48. package/dist/src/v2/contained-report-operation.d.ts.map +1 -1
  49. package/dist/src/v2/contained-report-operation.js +139 -466
  50. package/dist/src/v2/contained-report-operation.js.map +1 -1
  51. package/dist/src/v2/containment.d.ts +1 -0
  52. package/dist/src/v2/containment.d.ts.map +1 -1
  53. package/dist/src/v2/containment.js +12 -2
  54. package/dist/src/v2/containment.js.map +1 -1
  55. package/dist/src/v2/delivery-authority.d.ts +26 -0
  56. package/dist/src/v2/delivery-authority.d.ts.map +1 -0
  57. package/dist/src/v2/delivery-authority.js +44 -0
  58. package/dist/src/v2/delivery-authority.js.map +1 -0
  59. package/dist/src/v2/direct-delivery.d.ts +16 -36
  60. package/dist/src/v2/direct-delivery.d.ts.map +1 -1
  61. package/dist/src/v2/direct-delivery.js +135 -122
  62. package/dist/src/v2/direct-delivery.js.map +1 -1
  63. package/dist/src/v2/immutable-workflow-publisher.d.ts.map +1 -1
  64. package/dist/src/v2/immutable-workflow-publisher.js +3 -1
  65. package/dist/src/v2/immutable-workflow-publisher.js.map +1 -1
  66. package/dist/src/v2/implementation-report.d.ts +3 -1
  67. package/dist/src/v2/implementation-report.d.ts.map +1 -1
  68. package/dist/src/v2/implementation-report.js +17 -4
  69. package/dist/src/v2/implementation-report.js.map +1 -1
  70. package/dist/src/v2/implementation-reviewer.d.ts +41 -12
  71. package/dist/src/v2/implementation-reviewer.d.ts.map +1 -1
  72. package/dist/src/v2/implementation-reviewer.js +114 -42
  73. package/dist/src/v2/implementation-reviewer.js.map +1 -1
  74. package/dist/src/v2/pending-effect-settlement.d.ts +44 -0
  75. package/dist/src/v2/pending-effect-settlement.d.ts.map +1 -0
  76. package/dist/src/v2/pending-effect-settlement.js +69 -0
  77. package/dist/src/v2/pending-effect-settlement.js.map +1 -0
  78. package/dist/src/v2/process-identity.d.ts +45 -0
  79. package/dist/src/v2/process-identity.d.ts.map +1 -0
  80. package/dist/src/v2/process-identity.js +118 -0
  81. package/dist/src/v2/process-identity.js.map +1 -0
  82. package/dist/src/v2/proof-report.d.ts +2 -1
  83. package/dist/src/v2/proof-report.d.ts.map +1 -1
  84. package/dist/src/v2/proof-report.js +10 -4
  85. package/dist/src/v2/proof-report.js.map +1 -1
  86. package/dist/src/v2/review-feedback-coordinator.d.ts +1 -1
  87. package/dist/src/v2/review-feedback-coordinator.d.ts.map +1 -1
  88. package/dist/src/v2/review-feedback-coordinator.js +1 -1
  89. package/dist/src/v2/review-feedback-coordinator.js.map +1 -1
  90. package/dist/src/v2/review-feedback.d.ts +14 -20
  91. package/dist/src/v2/review-feedback.d.ts.map +1 -1
  92. package/dist/src/v2/review-feedback.js +45 -87
  93. package/dist/src/v2/review-feedback.js.map +1 -1
  94. package/dist/src/v2/run-issue.d.ts +129 -88
  95. package/dist/src/v2/run-issue.d.ts.map +1 -1
  96. package/dist/src/v2/run-issue.js +1965 -2381
  97. package/dist/src/v2/run-issue.js.map +1 -1
  98. package/dist/src/v2/run-state-projections.d.ts +84 -0
  99. package/dist/src/v2/run-state-projections.d.ts.map +1 -0
  100. package/dist/src/v2/run-state-projections.js +142 -0
  101. package/dist/src/v2/run-state-projections.js.map +1 -0
  102. package/dist/src/v2/run-store.d.ts +99 -81
  103. package/dist/src/v2/run-store.d.ts.map +1 -1
  104. package/dist/src/v2/run-store.js +245 -542
  105. package/dist/src/v2/run-store.js.map +1 -1
  106. package/dist/src/v2/runtime-assets.d.ts +3 -0
  107. package/dist/src/v2/runtime-assets.d.ts.map +1 -1
  108. package/dist/src/v2/runtime-assets.js +104 -0
  109. package/dist/src/v2/runtime-assets.js.map +1 -1
  110. package/dist/src/v2/runtime.d.ts +56 -44
  111. package/dist/src/v2/runtime.d.ts.map +1 -1
  112. package/dist/src/v2/runtime.js +383 -503
  113. package/dist/src/v2/runtime.js.map +1 -1
  114. package/dist/src/v2/setup.js +0 -2
  115. package/dist/src/v2/setup.js.map +1 -1
  116. package/dist/src/v2/validation-progression.d.ts +70 -0
  117. package/dist/src/v2/validation-progression.d.ts.map +1 -0
  118. package/dist/src/v2/validation-progression.js +247 -0
  119. package/dist/src/v2/validation-progression.js.map +1 -0
  120. package/dist/src/v2/workflow-assets.d.ts +9 -3
  121. package/dist/src/v2/workflow-assets.d.ts.map +1 -1
  122. package/dist/src/v2/workflow-assets.js +256 -43
  123. package/dist/src/v2/workflow-assets.js.map +1 -1
  124. package/internal-workflow/docs/agents/bug-workflow-routing.md +9 -7
  125. package/internal-workflow/docs/agents/coding-skill-routing.md +170 -120
  126. package/internal-workflow/docs/agents/tool-usage.md +23 -12
  127. package/internal-workflow/manifest.json +1 -1
  128. package/internal-workflow/operations/code-review/SKILL.md +34 -15
  129. package/internal-workflow/operations/implementation/SKILL.md +21 -16
  130. package/internal-workflow/profiles/implementer.toml +9 -0
  131. package/internal-workflow/profiles/review_coordinator.toml +9 -0
  132. package/internal-workflow/profiles/spec_reviewer.toml +9 -0
  133. package/internal-workflow/profiles/standards_reviewer.toml +9 -0
  134. package/internal-workflow/schemas/code-review-v1.json +1 -1
  135. package/internal-workflow/schemas/implementation-report-v1.json +1 -1
  136. package/internal-workflow/schemas/proof-report-v1.json +1 -1
  137. package/internal-workflow/skills/bug-root-cause-explainer/SKILL.md +114 -0
  138. package/internal-workflow/skills/bug-root-cause-explainer/agents/openai.yaml +7 -0
  139. package/internal-workflow/skills/bug-root-cause-explainer/evals/evals.json +18 -0
  140. package/internal-workflow/skills/code-review/SKILL.md +84 -306
  141. package/internal-workflow/skills/code-review/agents/openai.yaml +5 -3
  142. package/internal-workflow/skills/code-review/evals/evals.json +83 -0
  143. package/internal-workflow/skills/code-review/references/standards-smells.md +41 -0
  144. package/internal-workflow/skills/diagnosing-bugs/SKILL.md +69 -32
  145. package/internal-workflow/skills/diagnosing-bugs/agents/openai.yaml +2 -2
  146. package/internal-workflow/skills/diagnosing-bugs/evals/evals.json +63 -0
  147. package/internal-workflow/skills/grilling/SKILL.md +51 -0
  148. package/internal-workflow/skills/grilling/agents/openai.yaml +6 -0
  149. package/internal-workflow/skills/grilling/evals/evals.json +47 -0
  150. package/internal-workflow/skills/implement/SKILL.md +135 -0
  151. package/internal-workflow/skills/implement/agents/openai.yaml +6 -0
  152. package/internal-workflow/skills/implement/evals/evals.json +150 -0
  153. package/internal-workflow/skills/plan/SKILL.md +59 -0
  154. package/internal-workflow/skills/plan/agents/openai.yaml +6 -0
  155. package/internal-workflow/skills/plan/evals/evals.json +36 -0
  156. package/internal-workflow/skills/prototype/LOGIC.md +130 -0
  157. package/internal-workflow/skills/prototype/SKILL.md +69 -0
  158. package/internal-workflow/skills/prototype/UI.md +157 -0
  159. package/internal-workflow/skills/prototype/agents/openai.yaml +6 -0
  160. package/internal-workflow/skills/prototype/evals/evals.json +67 -0
  161. package/internal-workflow/skills/research/SKILL.md +110 -0
  162. package/internal-workflow/skills/research/agents/openai.yaml +6 -0
  163. package/internal-workflow/skills/research/evals/evals.json +49 -0
  164. package/internal-workflow/skills/tdd/SKILL.md +72 -67
  165. package/internal-workflow/skills/tdd/agents/openai.yaml +2 -2
  166. package/internal-workflow/skills/tdd/evals/evals.json +12 -0
  167. package/internal-workflow/skills/tdd/mocking.md +48 -1
  168. package/internal-workflow/skills/tdd/refactoring.md +3 -3
  169. package/internal-workflow/skills/tickets-orchestrator/SKILL.md +199 -0
  170. package/internal-workflow/skills/tickets-orchestrator/agents/openai.yaml +6 -0
  171. package/internal-workflow/skills/tickets-orchestrator/evals/evals.json +126 -0
  172. package/internal-workflow/skills/tickets-orchestrator/references/delegate-integrate.md +83 -0
  173. package/internal-workflow/skills/tickets-orchestrator/references/finish-delivery.md +69 -0
  174. package/internal-workflow/skills/tickets-orchestrator/references/stop-completion.md +63 -0
  175. package/internal-workflow/skills/to-spec/SKILL.md +133 -0
  176. package/internal-workflow/skills/to-spec/agents/openai.yaml +6 -0
  177. package/internal-workflow/skills/to-spec/evals/evals.json +24 -0
  178. package/internal-workflow/skills/to-tickets/SKILL.md +189 -0
  179. package/internal-workflow/skills/to-tickets/agents/openai.yaml +6 -0
  180. package/internal-workflow/skills/to-tickets/evals/evals.json +79 -0
  181. package/internal-workflow/skills/to-tickets/references/publishing-details.md +117 -0
  182. package/package.json +1 -1
  183. package/dist/src/v2/proof-store.d.ts +0 -54
  184. package/dist/src/v2/proof-store.d.ts.map +0 -1
  185. package/dist/src/v2/proof-store.js +0 -301
  186. package/dist/src/v2/proof-store.js.map +0 -1
  187. package/dist/src/v2/route-continuations.d.ts +0 -32
  188. package/dist/src/v2/route-continuations.d.ts.map +0 -1
  189. package/dist/src/v2/route-continuations.js +0 -2
  190. package/dist/src/v2/route-continuations.js.map +0 -1
  191. package/dist/src/v2/route-coordinator.d.ts +0 -72
  192. package/dist/src/v2/route-coordinator.d.ts.map +0 -1
  193. package/dist/src/v2/route-coordinator.js +0 -275
  194. package/dist/src/v2/route-coordinator.js.map +0 -1
  195. package/dist/src/v2/route-decision.d.ts +0 -120
  196. package/dist/src/v2/route-decision.d.ts.map +0 -1
  197. package/dist/src/v2/route-decision.js +0 -380
  198. package/dist/src/v2/route-decision.js.map +0 -1
  199. package/dist/src/v2/spec-coordinator.d.ts +0 -73
  200. package/dist/src/v2/spec-coordinator.d.ts.map +0 -1
  201. package/dist/src/v2/spec-coordinator.js +0 -126
  202. package/dist/src/v2/spec-coordinator.js.map +0 -1
  203. package/dist/src/v2/spec-delivery.d.ts +0 -112
  204. package/dist/src/v2/spec-delivery.d.ts.map +0 -1
  205. package/dist/src/v2/spec-delivery.js +0 -336
  206. package/dist/src/v2/spec-delivery.js.map +0 -1
  207. package/dist/src/v2/triage-route.d.ts +0 -68
  208. package/dist/src/v2/triage-route.d.ts.map +0 -1
  209. package/dist/src/v2/triage-route.js +0 -223
  210. package/dist/src/v2/triage-route.js.map +0 -1
  211. package/dist/src/v2/waiting-human-coordinator.d.ts +0 -49
  212. package/dist/src/v2/waiting-human-coordinator.d.ts.map +0 -1
  213. package/dist/src/v2/waiting-human-coordinator.js +0 -509
  214. package/dist/src/v2/waiting-human-coordinator.js.map +0 -1
  215. package/dist/src/v2/waiting-human.d.ts +0 -143
  216. package/dist/src/v2/waiting-human.d.ts.map +0 -1
  217. package/dist/src/v2/waiting-human.js +0 -408
  218. package/dist/src/v2/waiting-human.js.map +0 -1
  219. package/internal-workflow/docs/agents/contract-test-ledger.md +0 -71
  220. package/internal-workflow/docs/agents/review-gates.md +0 -42
  221. package/internal-workflow/docs/agents/review-protocol.md +0 -98
  222. package/internal-workflow/evals/coding-skill-evals.json +0 -373
  223. package/internal-workflow/operations/ambiguity-review/SKILL.md +0 -5
  224. package/internal-workflow/operations/qualification-repair/SKILL.md +0 -17
  225. package/internal-workflow/operations/spec-author/SKILL.md +0 -12
  226. package/internal-workflow/operations/spec-review/SKILL.md +0 -12
  227. package/internal-workflow/operations/triage/SKILL.md +0 -12
  228. package/internal-workflow/profiles/analyst_deep.toml +0 -9
  229. package/internal-workflow/profiles/implementer_standard.toml +0 -9
  230. package/internal-workflow/profiles/proof_agent.toml +0 -8
  231. package/internal-workflow/profiles/reviewer_deep.toml +0 -9
  232. package/internal-workflow/profiles/reviewer_standard.toml +0 -9
  233. package/internal-workflow/schemas/ambiguity-review-v1.json +0 -1
  234. package/internal-workflow/schemas/spec-author-v1.json +0 -1
  235. package/internal-workflow/schemas/spec-review-v1.json +0 -30
  236. package/internal-workflow/schemas/triage-route-v1.json +0 -1
  237. package/internal-workflow/skills/agent-auto/SKILL.md +0 -19
  238. package/internal-workflow/skills/agent-auto/agents/openai.yaml +0 -6
  239. package/internal-workflow/skills/code-debugger/SKILL.md +0 -122
  240. package/internal-workflow/skills/code-debugger/agents/openai.yaml +0 -7
  241. package/internal-workflow/skills/code-review/references/bug-classes.md +0 -56
  242. package/internal-workflow/skills/code-review/references/cleanup-lens.md +0 -52
  243. package/internal-workflow/skills/code-review/references/framework-lenses.md +0 -34
  244. package/internal-workflow/skills/code-review/references/targeted-recipes.md +0 -49
  245. package/internal-workflow/skills/implementation-spec-maker/SKILL.md +0 -107
  246. package/internal-workflow/skills/implementation-spec-maker/agents/openai.yaml +0 -6
  247. package/internal-workflow/skills/implementation-spec-maker/references/source-modes.md +0 -32
  248. package/internal-workflow/skills/implementation-spec-maker/references/spec-template.md +0 -146
  249. package/internal-workflow/skills/implementation-spec-review/SKILL.md +0 -131
  250. package/internal-workflow/skills/implementation-spec-review/agents/openai.yaml +0 -6
  251. package/internal-workflow/skills/implementation-spec-review/evals/evals.json +0 -78
  252. package/internal-workflow/skills/implementation-spec-review/references/review-loop.md +0 -121
  253. package/internal-workflow/skills/small-task-implementer/SKILL.md +0 -112
  254. package/internal-workflow/skills/small-task-implementer/agents/openai.yaml +0 -6
  255. package/internal-workflow/skills/spec-implementer/SKILL.md +0 -133
  256. package/internal-workflow/skills/spec-implementer/agents/openai.yaml +0 -6
  257. package/internal-workflow/skills/spec-implementer/evals/evals.json +0 -30
  258. package/internal-workflow/skills/spec-implementer/references/review-loop.md +0 -100
  259. package/internal-workflow/skills/triage/AGENT-BRIEF.md +0 -192
  260. package/internal-workflow/skills/triage/OUT-OF-SCOPE.md +0 -101
  261. package/internal-workflow/skills/triage/SKILL.md +0 -134
  262. package/internal-workflow/skills/triage/agents/openai.yaml +0 -6
@@ -1,32 +0,0 @@
1
- # Source Modes
2
-
3
- Read the section matching the active source. Apply the common source-authority and evidence rules from `SKILL.md` in every mode.
4
-
5
- ## Plan-Based
6
-
7
- - Treat the approved plan as architectural authority.
8
- - Preserve its scope, vertical-slice boundaries, guardrails, rejected paths, required docs, validation, and blocking assumptions.
9
- - Block when missing guardrails would let implementation drift or require redesign.
10
-
11
- ## Issue-Based
12
-
13
- - Treat one issue's acceptance criteria as the execution contract; parent material supplies product context, not sibling-ticket authority.
14
- - Read comments for changed decisions, blockers, credentials, external contracts, live prerequisites, and rejected approaches.
15
- - Preserve relevant `Implementation preparation`, `External contracts`, `Verification`, and `Blocked by` content.
16
- - Use `source_type: "issue"` when no plan exists. Do not block merely because `source_plan` is absent.
17
- - Block when acceptance criteria are ambiguous, non-verifiable, or contradicted by repository evidence.
18
-
19
- ## Contract Discovery
20
-
21
- - Specify discovery only: exact sources/tools to inspect, evidence to collect, decision record to update, and issue fields/comments to update.
22
- - Confirm the API surface, auth/secret source, license or terms constraints, acquisition path, deterministic fixture strategy, live-validation prerequisite, and rejected acquisition paths that matter to later implementation.
23
- - Do not include downstream implementation slices while material external behavior remains unconfirmed.
24
-
25
- ## Revision
26
-
27
- - Reconcile the entire existing spec against new authority and current repository evidence.
28
- - Treat a changed claim, mechanism/owner, source of truth, evidence unit/cardinality, or blocked/failure meaning as a semantic change; invalidate the affected approval coverage instead of translating it into broader wording.
29
- - Preserve still-valid completed `[x]` items exactly.
30
- - Reopen invalid completed items to `[ ]` and add a short `Revision Note:` with evidence.
31
- - Never silently delete progress or defect history.
32
- - Mark the spec blocked when completed history or its contract ledger cannot be trusted.
@@ -1,146 +0,0 @@
1
- # Implementation Spec Template
2
-
3
- Use the base template for every spec. Add conditional blocks only when their trigger applies, and remove every instruction or placeholder before review.
4
-
5
- ## Base Template
6
-
7
- ```markdown
8
- ---
9
- title: "<title>"
10
- created_at: "<ISO timestamp>"
11
- source_type: "plan | issue | contract-discovery | revised-spec"
12
- source_plan: "<absolute path or None>"
13
- source_issues:
14
- - "<URL/reference or None>"
15
- status: "draft | ready | blocked"
16
- execution_model: "single-agent | multi-agent"
17
- spec_mode: "compact | full"
18
- implementation_size: "small | medium | large"
19
- expected_repositories: <positive integer>
20
- review_profile: "simple | medium | high"
21
- review_reasons:
22
- - "<signal: evidence>"
23
- review_outcome: "Pending"
24
- review_verdict: "Not run"
25
- review_coverage: "Not reviewed"
26
- review_passes: "0"
27
- ---
28
-
29
- ## 1. Execution Context
30
- - **Goal:** <one observable outcome>
31
- - **Source Material:** <exact references>
32
- - **Approved Scope:** <strict allowed work>
33
- - **Out of Scope:** <explicit exclusions or None>
34
- - **Minimum Solution:** <direct path through existing owners and public seams>
35
- - **Added Complexity:** None | <repeat one entry per mechanism: `<mechanism>` — required for `<invariant or evidenced failure>`; without it `<concrete breakage>`>
36
- - **Primary Risk:** <main correctness or coordination risk>
37
-
38
- ## 2. Preconditions And Evidence
39
- - **Required Services / Env / Fixtures:** <exact requirements or None>
40
- - **Blocking Unknowns:** <exact unknowns when blocked, otherwise None>
41
- - **Confirmed Targets:** <minimal evidence-backed paths and symbols>
42
- - **Confirmed Commands:** <exact commands>
43
- - **Protected Paths / Rejected Approaches:** <items or None>
44
- - **Source of Truth:** <existing owner whenever behavior/data can drift; otherwise omit>
45
- - **New Boundaries:** <only when ownership or a public seam changes; otherwise omit>
46
-
47
- ## 3. Execution Slices
48
-
49
- ### Slice 1 — <narrow end-to-end behavior>
50
- - [ ] **Test/Proof First:** <failing behavior test or exact observable proof>
51
- - [ ] **Target:** `<exact/path:symbol>` — <specific action>
52
- - [ ] **Validation:** <target-level check>
53
- - [ ] **Exit Gate:** <command or proof that the slice works end-to-end>
54
-
55
- <repeat only for independently verifiable behavior slices>
56
-
57
- ## 4. Validation And Done Criteria
58
- - [ ] **Lint/Format:** <exact command or Not applicable with reason>
59
- - [ ] **Typecheck/Build:** <exact command or Not applicable with reason>
60
- - [ ] **Tests:** <exact command or Not applicable with reason>
61
- - [ ] **Architecture Check:** <exact command or Not applicable with reason>
62
- - [ ] **Live/Manual Proof:** <exact flow or Not applicable with reason>
63
- - [ ] **Behavior Proof:** <observable acceptance proof>
64
- - [ ] **Reconciliation:** every unchecked item is unfinished, blocked with evidence, or intentionally not applicable.
65
- - [ ] **Final Handoff Requirements:** <high only: extended `$spec-implementer` Final Risk Handoff plus task-specific deviations; omit for ordinary medium work>
66
- ```
67
-
68
- ## Conditional Blocks
69
-
70
- ### Contract Test Ledger
71
-
72
- Add for contract-heavy behavior using the shared Contract Test Ledger referenced by `SKILL.md`. Keep one row per material invariant and place it before execution slices.
73
-
74
- ### Review Checkpoint And Focus
75
-
76
- Add a checkpoint only when the risky target becomes stable before later slices. Otherwise put this compact block in final review coverage:
77
-
78
- ```markdown
79
- ## Review Focus
80
- - **Mandatory Lenses:** <applicable lenses>
81
- - **Targeted Recipes:** <applicable recipes or None>
82
- - **Bug Classes:** <concrete failures to hunt>
83
- ```
84
-
85
- ### Risk Controls
86
-
87
- Add in `full` mode only for applicable ambiguity:
88
-
89
- ```markdown
90
- ## Risk Controls
91
- - **Source of Truth:** <owner>
92
- - **Safety / Contract / State Constraints:** <only applicable constraints>
93
- - **Forbidden Scope:** <tempting but rejected paths>
94
- - **Review Timing:** <stable early checkpoint or concrete final-review focus>
95
- ```
96
-
97
- ### Write Scope Summary
98
-
99
- Add for multi-agent work, generated artifacts, broad runtime changes, or when phase targets do not make the write set auditable.
100
-
101
- ```markdown
102
- ## Write Scope Summary
103
- - `<path>` — <Create | Update | Delete>; <responsibility>
104
- ```
105
-
106
- ### Integrator Coordination Contract
107
-
108
- Require only when `execution_model: "multi-agent"`:
109
-
110
- ```markdown
111
- ## Integrator Coordination Contract
112
- | Agent | Exclusive Write Scope | Handoff | Merge Phase |
113
- | --- | --- | --- | --- |
114
- | <agent> | <disjoint paths> | <artifact> | <order> |
115
-
116
- - **Integrator Owner:** <owner>
117
- - **Forbidden Overlap:** <paths/contracts>
118
- - **Final Duties:** <integration, validation, reconciliation>
119
- ```
120
-
121
- ### Halt Conditions
122
-
123
- Add only when the common contradiction/guessing stop rule is insufficient. Use 3–6 task-specific conditions.
124
-
125
- ### Defect Closure Notes
126
-
127
- Add only when review returns defects:
128
-
129
- ```markdown
130
- ## Defect Closure Notes
131
- - **Review Summary:** <pass counts and coverage>
132
- - **Verified Defects:** <stable IDs or None>
133
- - **Accepted Risks:** <stable IDs, authority, and reason or None>
134
- - **Open Defects:** <stable IDs or None>
135
- ```
136
-
137
- ## Terminal Review Metadata
138
-
139
- Replace the temporary frontmatter values after the owner review loop returns a real terminal outcome:
140
-
141
- ```yaml
142
- review_outcome: "Approved | Blocked | Waived"
143
- review_verdict: "Approved | Needs Work | Rejected | Not run"
144
- review_coverage: "<covered lenses or Not reviewed>"
145
- review_passes: "<total; full/closure/fresh counts>"
146
- ```
@@ -1,131 +0,0 @@
1
- ---
2
- name: "implementation-spec-review"
3
- description: "Review compact or full implementation specs for deterministic executability, proportional scope, validation coverage, safety, and zero-guess execution before coding starts."
4
- ---
5
-
6
- # Implementation Spec Review
7
-
8
- Decide whether a saved implementation spec can be executed safely without
9
- guessing. Review execution quality, not the product idea. Do not rewrite the
10
- spec unless explicitly asked.
11
-
12
- Read:
13
-
14
- - `references/review-loop.md` when called by `$implementation-spec-maker`;
15
- - `../../docs/agents/confidence-rubric.md` for defect confidence;
16
- - `../../docs/agents/contract-test-ledger.md` only when the spec changes a
17
- material behavior contract.
18
-
19
- ## Independent Dimensions
20
-
21
- Keep these classifications independent:
22
-
23
- - `spec_mode: compact | full` — document/coordination density;
24
- - `implementation_size: small | medium | large` — delivery shape;
25
- - `review_profile: simple | medium | high` — consequence and uncertainty;
26
- - `expected_repositories` — approved repository count.
27
-
28
- Compact may describe broad or high-risk work when ownership, sequencing, and
29
- proof remain deterministic. Full is justified only when concrete coordination,
30
- contract, safety, ownership, or validation ambiguity cannot fit clearly in the
31
- compact form. Never request full-mode tables or ceremony merely from size or
32
- risk labels.
33
-
34
- ## Adapter Contract
35
-
36
- When called by the maker, use the mode and lenses supplied by
37
- `references/review-loop.md`, reuse supplied defect IDs, and return actual
38
- coverage. A reviewer child executes this Adapter inline and never spawns a
39
- grandchild. If root receives a direct review request, it launches the
40
- profile-selected reviewer instead of self-reviewing.
41
-
42
- A standalone reviewer performs one bounded Full over all applicable lenses and
43
- returns only `Approved | Needs Work | Rejected`; it does not invent owner state
44
- or claim Closure.
45
-
46
- ## Minimum Solution First
47
-
48
- Begin every Full review with the evidence-backed scope delta, before checking whether the proposed implementation is detailed enough:
49
-
50
- 1. Identify behavior already implemented and require preservation/regression proof instead of reimplementation.
51
- 2. Challenge every new endpoint, service, durable state, configuration input, schema/public contract, repository, and data owner against an existing owner or seam.
52
- 3. Require one approved requirement or concrete failure path for each surviving mechanism. If deletion still satisfies behavior, invariants, and proof, report the mechanism as excess.
53
- 4. Treat invented product values, copy, eligibility policy, thresholds, and defaults as authority gaps when they shape observable behavior; detailed implementation does not resolve them.
54
-
55
- Prefer the smallest repair in this order: delete excess, reuse an existing owner/seam, narrowly extend an existing contract, then add a new mechanism only when the earlier options cannot satisfy a named invariant.
56
-
57
- ## Review Lenses
58
-
59
- Scale depth to the profile and inspect only applicable lenses:
60
-
61
- - **Determinism and evidence:** execution-critical paths, symbols, commands,
62
- contracts, fixtures, and claims are confirmed rather than invented.
63
- - **Scope and minimum solution:** the spec preserves approved scope,
64
- distinguishes `preserve + regression proof` from new work, uses existing
65
- owners/seams, and ties every added mechanism to a requirement or concrete
66
- failure path.
67
- - **Sequencing and ownership:** phases are safe, sources of truth are explicit
68
- where drift is possible, and multi-agent write scopes are disjoint.
69
- - **Validation:** each behavior has an observable proof; contract-risk work maps
70
- each material invariant to its first failing test or exact blocked proof.
71
- - **Preconditions and stop conditions:** required services, data, env, fixtures,
72
- and destructive/sensitive constraints are explicit when applicable.
73
- - **Review focus:** ordinary work relies on one final review; only an explicit
74
- stable high-risk slice gets an intermediate checkpoint.
75
- - **Revision integrity:** current content matches its authority and preserves
76
- still-valid completed work.
77
- - **Completion:** another agent can tell what to do, what proves success, when
78
- to stop, and what remains blocked.
79
-
80
- Before accepting the supplied scope delta, independently recover each material claim from the raw authority and compare its mechanism/owner, source of truth, evidence unit/cardinality, and blocked/failure meaning. Report compressed or missing decisions. Challenge each material proof with: can this source observe the exact claim, and can the gate pass while the claim is false? Shared proofs across platforms, providers, tenants, regions, or modes require evidenced equivalence in mechanism, truth source, granularity, timing/redaction, and failure semantics.
81
-
82
- ## Proportional Expectations
83
-
84
- Approve a compact spec when targets, ordered work, observable proof, and stop
85
- conditions are exact enough for the task. Do not require source-of-truth tables,
86
- file matrices, long halt lists, multi-agent contracts, or defect sections when
87
- no concrete ambiguity needs them.
88
-
89
- A lean full spec normally adds only applicable `Risk Controls`, exact phase
90
- targets/proof, and—when needed—write-scope or integrator coordination. Missing
91
- ownership, validation, safety, or handoff detail is a defect; missing formatting
92
- ceremony is not.
93
-
94
- Prefer deleting or narrowing an unsafe proposal before adding flags, telemetry,
95
- fallbacks, compatibility paths, or rollout machinery. Optional improvements
96
- remain optional unless source authority approves them.
97
-
98
- Do not approve a duplicate public seam merely because it is internally consistent. Do not ask for more detailed recovery, configuration, analytics, or compatibility machinery until the mechanism itself passes the deletion challenge.
99
-
100
- ## Defects And Decision
101
-
102
- - **Blocker:** unsafe or impossible to execute as written.
103
- - **Execution risk:** executable but likely to drift or require rework.
104
- - **Improvement:** useful but not required for safe execution.
105
-
106
- Reject exact-looking but ungrounded paths/contracts, unresolved placeholders or
107
- alternative commands, validation that cannot prove the intended behavior,
108
- overlapping multi-agent ownership, missing material safety constraints, or any
109
- step that requires invention.
110
-
111
- Use:
112
-
113
- - `Approved` when the current spec is deterministic, bounded, proportional, and
114
- executable without guessing;
115
- - `Needs Work` for repairable ambiguity, weak proof, or excess ceremony;
116
- - `Rejected` when execution would be unsafe or depend on invented decisions.
117
-
118
- ## Output
119
-
120
- Answer in Russian and keep technical terms in English. Return:
121
-
122
- 1. `Вердикт` and one-sentence reason.
123
- 2. `Режим и покрытие` with Full/Closure and actual lenses.
124
- 3. Short `Determinism / Evidence / Validation / Safety` scores from 0 to 2.
125
- 4. Evidence-backed defects first, with supplied ID or `NEW-<LENS>-NN`, class,
126
- confidence, failure, evidence, smallest repair, and affected section.
127
- 5. Exact changes needed before execution, or `Ничего`.
128
- 6. Only genuinely blocking questions, or `Нет`.
129
-
130
- Do not repeat the spec, propose broad redesign, or turn optional cleanup into a
131
- mandatory gate.
@@ -1,6 +0,0 @@
1
- interface:
2
- display_name: "Implementation Spec Review"
3
- short_description: "Review one implementation spec"
4
- default_prompt: "Review the supplied spec through the assigned package-owned operation and return its exact JSON report."
5
- policy:
6
- allow_implicit_invocation: false
@@ -1,78 +0,0 @@
1
- {
2
- "schema_version": 1,
3
- "skill": "implementation-spec-review",
4
- "cases": [
5
- {
6
- "id": "artifact-profile-by-consequence",
7
- "prompt": "Review one broad but reversible spec and one narrow spec with irreversible data impact and unclear recovery ownership.",
8
- "expected": ["broad reversible may remain medium", "narrow dangerous uncertain spec is high"],
9
- "forbidden": ["classify from file count or spec mode"]
10
- },
11
- {
12
- "id": "artifact-scope-conservation",
13
- "prompt": "A review can repair the spec either by deleting an unnecessary mechanism or by adding flags, telemetry, and fallback infrastructure.",
14
- "expected": ["prefer the smallest scope-preserving repair"],
15
- "forbidden": ["add unapproved operational machinery"]
16
- },
17
- {
18
- "id": "artifact-approval-invalidation",
19
- "prompt": "An approved spec receives a substantive execution change after review.",
20
- "expected": ["invalidate approval for the changed revision", "review only invalidated coverage"],
21
- "forbidden": ["execute under stale approval", "restart unrelated coverage"]
22
- },
23
- {
24
- "id": "artifact-existing-seam-extension",
25
- "prompt": "A spec proposes a second public access endpoint and new client loader although the existing versioned endpoint can accept optional fields without breaking callers.",
26
- "expected": ["report excess scope", "prefer a narrow extension of the existing endpoint and loader"],
27
- "forbidden": ["approve the duplicate seam because its contract is detailed", "add compatibility machinery for the duplicate endpoint"]
28
- },
29
- {
30
- "id": "artifact-already-implemented-preservation",
31
- "prompt": "Repository evidence proves the target screen already displays the current plan and remaining time, but the spec lists both as new implementation work.",
32
- "expected": ["reclassify existing behavior as preserve plus regression proof", "limit implementation to the actual remaining delta"],
33
- "forbidden": ["plan a second implementation", "ignore current code evidence"]
34
- },
35
- {
36
- "id": "artifact-invented-product-policy",
37
- "prompt": "The source requires configurable lookback and cooldown periods but supplies no values; the spec invents defaults and builds configuration, validation, and analytics around them.",
38
- "expected": ["treat material values as an authority gap", "block before expanding the technical contract"],
39
- "forbidden": ["approve invented defaults", "reward extra implementation detail as determinism"]
40
- },
41
- {
42
- "id": "artifact-closure-complexity-guard",
43
- "prompt": "A repair for one failure-contract defect adds a service, durable state machine, configuration input, and public DTO. Review the repair in Closure.",
44
- "expected": ["verify the supplied defect", "challenge each added mechanism against existing seams", "report scope change without starting a separate simplification review"],
45
- "forbidden": ["verify only the old defect ID", "automatically launch a fresh Full"]
46
- },
47
- {
48
- "id": "artifact-local-repair-review-budget",
49
- "prompt": "A consolidated repair changes only wording and one existing validation command while preserving scope, owners, public contracts, and mandatory-lens coverage.",
50
- "expected": ["use coordinator verification for ordinary findings", "reuse valid Full coverage"],
51
- "forbidden": ["launch Closure for ordinary findings", "launch a fresh Full because the revision changed"]
52
- },
53
- {
54
- "id": "artifact-authority-compression",
55
- "prompt": "Raw authority replaces a per-item provider mechanism with aggregate reporting, but the revised spec calls it only privacy-preserving processing and keeps per-item timestamp proof.",
56
- "expected": ["recover the concrete authority change", "reject the incompatible proof cardinality", "require a fresh Full"],
57
- "forbidden": ["trust the compressed scope delta", "approve a per-item proof for an aggregate claim"]
58
- },
59
- {
60
- "id": "artifact-variant-false-symmetry",
61
- "prompt": "Two platforms share one product goal but use different mechanisms, sources of truth, reporting delays, and redaction rules; the spec gives them one identical exit gate.",
62
- "expected": ["split the variant proof contracts unless equivalence is evidenced", "test whether each source can observe its exact claim"],
63
- "forbidden": ["approve symmetry from the shared product goal"]
64
- },
65
- {
66
- "id": "artifact-closure-semantic-change",
67
- "prompt": "During Closure a repair changes the source of truth and proof from individual records to an aggregate report while leaving the implementation owner unchanged.",
68
- "expected": ["invalidate affected coverage", "start a fresh Full for the semantic change"],
69
- "forbidden": ["verify only the old defect", "treat unchanged code ownership as sufficient"]
70
- },
71
- {
72
- "id": "artifact-proof-only-repair",
73
- "prompt": "Full review finds aggregate proof for an unchanged per-item claim; the repair substitutes an existing compatible per-item proof without changing authority, claim, owner, or source of truth.",
74
- "expected": ["reuse valid Full coverage", "verify the affected proof through coordinator verification or normal Closure triggers"],
75
- "forbidden": ["start a fresh Full for the proof correction"]
76
- }
77
- ]
78
- }
@@ -1,121 +0,0 @@
1
- # Implementation Spec Review Loop
2
-
3
- This reference owns review orchestration for specs created by
4
- `implementation-spec-maker`. Read it when the maker requests artifact review.
5
- The reviewer skill remains the Adapter; the shared review mechanics live in
6
- `../../../docs/agents/review-protocol.md`.
7
-
8
- ## Contract
9
-
10
- Input:
11
-
12
- - saved spec and pinned revision;
13
- - source authority and approved decisions;
14
- - evidence needed to verify execution claims;
15
- - optional user-raised review profile.
16
-
17
- Output:
18
-
19
- - `outcome: Approved | Blocked | Waived`;
20
- - `adapter_verdict: Approved | Needs Work | Rejected | Not run`;
21
- - `review_profile: simple | medium | high` and evidence-backed reasons;
22
- - mandatory-lens coverage and unresolved defects.
23
-
24
- The Adapter returns only its verdict. The root maps preflight, convergence, and
25
- waiver state to the artifact outcome.
26
-
27
- ## Preflight And Profile
28
-
29
- Before launching a reviewer, confirm source authority, approved scope, current
30
- spec revision, and mandatory external evidence. Save a useful blocked spec when
31
- a product or contract decision is missing; do not launch review to discover a
32
- known authority gap.
33
-
34
- Require an evidence-backed scope delta before review: approved behavior, current capability/owner/seam, and the smallest remaining implementation delta for each material requirement. Already implemented behavior is preservation/regression scope. Unresolved product values, copy, policy, thresholds, or ownership choices that materially shape behavior block preflight instead of receiving invented defaults.
35
-
36
- `medium` is the default. Use:
37
-
38
- - `simple` for one narrow owner with direct proof and no material uncertainty;
39
- - `medium` for all ordinary specs, including multi-file, API, persistence, or
40
- stateful work with clear ownership and bounded proof;
41
- - `high` only when a sensitive mechanism has both a material failure
42
- consequence and an uncertainty amplifier such as unclear ownership,
43
- cross-trust effects, non-local recovery, or an unproven external contract.
44
-
45
- File count and implementation size never select `high`. The user may raise but
46
- not lower an evidence-backed profile.
47
-
48
- ## Scope And Capsule
49
-
50
- Review the smallest approved solution. Risk may strengthen proof but does not
51
- authorize flags, telemetry, compatibility paths, generic fallbacks, or rollout
52
- machinery unless the source or a concrete failure requires them.
53
-
54
- Give each reviewer a bounded capsule containing the current spec, authority,
55
- approved scope, scope delta, current owners/seams, evidence, review question,
56
- assigned lenses, and current defect records. Name behavior that already exists
57
- and the justification for every proposed new endpoint, service, durable state,
58
- configuration input, schema/public contract, repository, or data owner. For
59
- Closure also include the repaired sections, affected contracts, and repair
60
- complexity delta. Do not pass raw parent history or unrelated inventories.
61
-
62
- ## Topology
63
-
64
- - `simple`: one `reviewer_fast`, one bounded Full.
65
- - `medium`: one `reviewer_standard`, one bounded Full.
66
- - `high`: two parallel `reviewer_deep` sessions with disjoint primary lenses:
67
- Architecture/Execution and Failure/Contracts.
68
-
69
- Root launches and aggregates reviewers. A reviewer child runs the
70
- `implementation-spec-review` Adapter inline and never spawns another reviewer.
71
- Reuse valid coverage for the same revision and question. Parallel high-profile
72
- reviewers are one Full review round, not sequential rounds.
73
-
74
- After one consolidated repair, coordinator verification is enough for ordinary
75
- medium/low findings. Use shared-protocol Closure only for critical/high defects,
76
- protected trust/data/concurrency/shared-contract impact, or invalidated
77
- mandatory coverage.
78
-
79
- The default budget is one Full round, one consolidated repair, and at most one
80
- Closure round. Closure verifies its supplied defects and runs a bounded
81
- complexity guard over the repair delta:
82
-
83
- 1. Did the repair add a new integration boundary or durable mechanism?
84
- 2. Can it be removed or replaced by an existing owner/seam?
85
- 3. Did it change approved scope?
86
-
87
- This guard is part of Closure, not a separate simplification review. A further
88
- targeted Closure is exceptional and requires a newly introduced critical/high
89
- defect plus materially changed target or evidence; otherwise coordinator
90
- verification or the shared no-progress/blocked outcome applies.
91
-
92
- Start a fresh Full only when an authority or approved-contract change
93
- invalidates existing mandatory coverage by changing the observable claim,
94
- primary mechanism/owner, source of truth, required evidence unit/cardinality,
95
- variant-equivalence assumption, blocked/failure meaning, approved scope,
96
- repository/data owner, public API, or durable workflow. Closure cannot absorb
97
- such a semantic change. A repair that only replaces incompatible proof with
98
- proof matching an unchanged reviewed claim uses coordinator verification or
99
- Closure under the normal triggers; wording, command, and other local repairs
100
- also reuse valid coverage.
101
-
102
- ## Approval
103
-
104
- Return `Approved` only when the current saved revision matches source authority,
105
- mandatory lenses are covered, and every blocking defect is verified. Any
106
- substantive edit invalidates approval; lifecycle metadata alone does not.
107
-
108
- Return `Blocked` when authority/evidence is missing, repair needs a product or
109
- ownership decision, no substantive repair exists, or shared no-progress rules
110
- apply. Return `Waived` only after explicit user instruction and keep skipped
111
- coverage visible; an open blocker still maps the artifact to `Blocked`.
112
-
113
- Map outcomes to spec status:
114
-
115
- - `Approved` -> `ready`;
116
- - `Blocked` -> `blocked`;
117
- - eligible `Waived` -> `ready` with visible waiver metadata.
118
-
119
- Report profile, outcome, Adapter verdict, mandatory coverage, verified/open
120
- defects, and skipped checks. Do not report counters or session history for a
121
- normal one-review flow.
@@ -1,112 +0,0 @@
1
- ---
2
- name: small-task-implementer
3
- description: Implement small low-risk coding tasks with narrow edits and targeted validation. Use for tiny local fixes, UI/copy changes, config/build corrections, or simple tests that do not change a shared API, DTO, schema, persistence, auth, or other review-gated contract.
4
- ---
5
-
6
- # Small Task Implementer
7
-
8
- Use this skill for fast, bounded implementation when creating a PRD and approved ticket-delivery flow would be disproportionate.
9
-
10
- ## Fit Gate
11
-
12
- Proceed only when all are true:
13
-
14
- - The requested behavior is clear or can be inferred from local code/tests without a product decision.
15
- - The change is expected to touch one small area or a few tightly related files.
16
- - There is a narrow validation path: targeted test, lint/typecheck, build check, UI proof, or direct command.
17
- - The task does not require a new plan, PRD, issue breakdown, implementation spec, migration, rollout, or multi-agent orchestration.
18
- - The change does not alter a shared API, DTO, schema, persistence, auth,
19
- permission, payment, cache, concurrency, background-job, navigation, or other
20
- contract covered by `docs/agents/review-gates.md`.
21
-
22
- For a tiny local behavior change that still fits this gate, `$tdd` is an
23
- independent proof method: activate it when the global TDD Fit Gate passes. Do
24
- not relabel a feature as a bug or use `$code-debugger` unless repository
25
- evidence confirms a defect.
26
-
27
- Escalate out of the tiny-task route when the work has more than one coherent
28
- behavior, a broad ownership boundary, material rollback/recovery risk, unclear
29
- product intent, no credible affected validation, or genuine multi-agent/live
30
- coordination. Statefulness alone does not require escalation into planning or
31
- orchestration; a clear API, persistence, cache, queue, or DTO change normally
32
- becomes direct medium implementation.
33
-
34
- Escalation rule:
35
-
36
- - For clear authority and one coherent outcome, escalate to direct medium root
37
- implementation under `$tdd`, affected validation, and one final review when
38
- `review-gates.md` applies.
39
- - Use optional `$grilling`, then `$spec-to-tickets` and `$tickets-orchestrator`,
40
- only for unresolved product decisions or a real approved ticket graph,
41
- delivery dependency, or explicit orchestration request.
42
- - For one risky behavior or technical contract, prefer one approved ticket and mark `compact spec` or `standard spec` only when the ticket plus repository evidence cannot remove execution ambiguity.
43
- - For several tickets sharing one unresolved contract or validation path, make the contract-defining ticket block its consumers; merge tickets that cannot be specified or verified independently instead of creating a wave-level implementation spec.
44
- - Escalate if the bug requires Bugfix Quality Gate analysis across multiple paths, states, async events, persistence, auth, cache, retries, workers, or contracts.
45
-
46
- ## Workflow
47
-
48
- 1. Inspect local context just enough to confirm fit.
49
- - Read repo instructions and the smallest relevant code/test files.
50
- - Check `git status --short` before editing.
51
- - Preserve unrelated dirty work.
52
-
53
- 2. Write a compact contract in the working update or internal task notes:
54
-
55
- ```text
56
- Behavior:
57
- Scope boundary:
58
- Validation:
59
- ```
60
-
61
- 3. Implement the smallest complete change.
62
- - Prefer existing patterns and owner modules.
63
- - Avoid unrelated refactors, abstractions, cleanup, and compatibility paths.
64
- - Before adding a helper, module, layer, or seam, apply the deletion test; keep it only if it improves current locality or leverage.
65
- - Do not add pass-through modules, one-adapter seams, or tests coupled to Implementation details; escalate if no natural public test seam exists.
66
- - Add or update a focused test only when behavior risk justifies it and the repo has a natural seam.
67
- - For pure copy/docs/config changes, do not invent tests; run the cheapest relevant syntax/lint/check instead.
68
-
69
- 4. Run targeted validation.
70
- - Use the narrowest meaningful command first.
71
- - If validation is unavailable or too expensive, state the concrete reason and residual risk.
72
- - Do not run full CI unless local policy or the changed surface makes it necessary.
73
-
74
- 5. Stop and escalate if implementation reveals hidden risk.
75
- - Examples: shared contract drift, duplicate source of truth, missing test
76
- seam, material concurrency/recovery uncertainty, or product ambiguity.
77
- - Leave a short explanation of what was discovered and which heavier flow should take over.
78
-
79
- ## Output
80
-
81
- Final response must stay compact:
82
-
83
- ```text
84
- Small Task Result
85
-
86
- Changed:
87
- - ...
88
-
89
- Proof:
90
- - ...
91
-
92
- Skipped:
93
- - none / ...
94
-
95
- Risk:
96
- - low / reason
97
- ```
98
-
99
- If escalated, use:
100
-
101
- ```text
102
- Escalated
103
-
104
- Reason:
105
- - ...
106
-
107
- Recommended flow:
108
- - Direct medium implementation / canonical ticket delivery
109
-
110
- Evidence:
111
- - ...
112
- ```
@@ -1,6 +0,0 @@
1
- interface:
2
- display_name: "Small Task Implementer"
3
- short_description: "Fast bounded local low-risk code changes"
4
- default_prompt: "Use $small-task-implementer to make this small local low-risk change with targeted validation, escalating shared contracts to the direct medium route."
5
- policy:
6
- allow_implicit_invocation: true