@tea-agent/loop-agent 0.13.0 → 0.15.0

This diff represents the content of publicly available package versions that have been released to one of the supported registries. The information contained in this diff is provided for informational purposes only and reflects changes between package versions as they appear in their respective public registries.
Files changed (272) hide show
  1. package/AGENTS.md +157 -157
  2. package/CHANGELOG.md +116 -305
  3. package/README.md +357 -334
  4. package/bin/agent-worker.js +22 -22
  5. package/bin/loop-agent.js +21 -21
  6. package/dist/commands/cursor-prompt.js +6 -6
  7. package/dist/commands/init.js +505 -505
  8. package/dist/commands/loop-benchmark.js +11 -11
  9. package/dist/commands/pi-reuse-benchmark.js +16 -16
  10. package/dist/executors/pi-event-serializer.js +33 -11
  11. package/dist/sidecars/cursor-prompt/executor.js +1 -1
  12. package/dist/task/runtime.js +27 -27
  13. package/dist/worker/observe/spec-evidence.js +19 -10
  14. package/dist/worker/observe/static/api.js +46 -46
  15. package/dist/worker/observe/static/app.js +151 -150
  16. package/dist/worker/observe/static/constants.js +156 -148
  17. package/dist/worker/observe/static/copy.js +67 -67
  18. package/dist/worker/observe/static/dag-helpers.js +201 -172
  19. package/dist/worker/observe/static/dag-layout.d.ts +31 -31
  20. package/dist/worker/observe/static/dag-layout.js +83 -83
  21. package/dist/worker/observe/static/dag-model.js +72 -72
  22. package/dist/worker/observe/static/dom.js +122 -122
  23. package/dist/worker/observe/static/format-pool.d.ts +71 -0
  24. package/dist/worker/observe/static/format-pool.js +134 -67
  25. package/dist/worker/observe/static/format.js +317 -292
  26. package/dist/worker/observe/static/index.html +350 -308
  27. package/dist/worker/observe/static/kpi.js +100 -94
  28. package/dist/worker/observe/static/markdown-render.js +124 -0
  29. package/dist/worker/observe/static/relations.js +133 -133
  30. package/dist/worker/observe/static/router.js +93 -93
  31. package/dist/worker/observe/static/run-processing.js +148 -148
  32. package/dist/worker/observe/static/shell-chrome.js +74 -68
  33. package/dist/worker/observe/static/state.js +273 -267
  34. package/dist/worker/observe/static/styles.css +2504 -1902
  35. package/dist/worker/observe/static/views/batch.js +227 -227
  36. package/dist/worker/observe/static/views/dag-graph.js +172 -172
  37. package/dist/worker/observe/static/views/dag-inspector.js +530 -627
  38. package/dist/worker/observe/static/views/dag.js +371 -371
  39. package/dist/worker/observe/static/views/dashboard.js +86 -100
  40. package/dist/worker/observe/static/views/failures.js +143 -143
  41. package/dist/worker/observe/static/views/feature.js +492 -492
  42. package/dist/worker/observe/static/views/pool.js +708 -350
  43. package/dist/worker/observe/static/views/run.js +453 -453
  44. package/dist/worker/observe/static/views/session-timeline.js +771 -219
  45. package/dist/worker/observe/static/views/shell.js +7 -7
  46. package/dist/worker/observe/static/views/task.js +314 -314
  47. package/dist/worker/observe/static/views/timeline.js +163 -163
  48. package/dist/workflows/dag/canvas-observer.js +275 -275
  49. package/dist/workflows/dag/init-hybrid.js +27 -11
  50. package/docs/README.md +106 -104
  51. package/docs/architecture/README.md +26 -26
  52. package/docs/architecture/dag-execution.md +140 -140
  53. package/docs/architecture/evolution.md +54 -54
  54. package/docs/architecture/facts-and-state.md +71 -71
  55. package/docs/architecture/runtime-boundaries.md +191 -191
  56. package/docs/architecture/system-overview.md +93 -93
  57. package/docs/architecture/worker-and-feature.md +85 -85
  58. package/docs/harness-methodology-debugging.md +153 -153
  59. package/docs/harness-methodology-tdd.md +130 -130
  60. package/docs/harness-methodology-verification.md +27 -27
  61. package/docs/init-surface.manifest.json +304 -307
  62. package/docs/skills/README.md +7 -7
  63. package/docs/skills/vetted-skill-registry.md +29 -29
  64. package/docs/templates/adr.md +60 -60
  65. package/docs/templates/agent-dag-authority-surface-audit.prompt.md +94 -94
  66. package/docs/templates/agent-dag-decision-envelope.schema.json +213 -213
  67. package/docs/templates/agent-dag-decision-gate-dogfood-report.md +117 -117
  68. package/docs/templates/agent-dag-decision-gate.prompt.md +246 -246
  69. package/docs/templates/agent-dag-process-supervisor.prompt.md +98 -98
  70. package/docs/templates/agent-dag-report.schema.json +473 -473
  71. package/docs/templates/agent-dag-review-verdict.prompt.md +68 -68
  72. package/docs/templates/agent-dag.base.json +190 -190
  73. package/docs/templates/agent-dag.final-verification.json +185 -185
  74. package/docs/templates/agent-dag.schema.json +411 -411
  75. package/docs/templates/agent-dag.supervised-implementation.json +620 -620
  76. package/docs/templates/backend-test-analysis.schema.json +44 -44
  77. package/docs/templates/backend-test-case-manifest.schema.json +190 -190
  78. package/docs/templates/backend-test-dag.classify.prompt.md +75 -75
  79. package/docs/templates/backend-test-dag.generate-pytest.prompt.md +204 -204
  80. package/docs/templates/backend-test-dag.json +559 -559
  81. package/docs/templates/backend-test-dag.retrospect.prompt.md +139 -139
  82. package/docs/templates/backend-test-dag.review-cases.prompt.md +83 -83
  83. package/docs/templates/backend-test-execution.schema.json +133 -133
  84. package/docs/templates/backend-test-result.schema.json +99 -99
  85. package/docs/templates/branch-merge-report.md +0 -1
  86. package/docs/templates/exec-plan.md +64 -64
  87. package/docs/templates/feature-spec.md +53 -53
  88. package/docs/templates/frontend-design-contract.md +42 -42
  89. package/docs/templates/frontend-eval/fixtures/failures/01-type-build-error.md +17 -17
  90. package/docs/templates/frontend-eval/fixtures/failures/02-unit-component-test-fail.md +16 -16
  91. package/docs/templates/frontend-eval/fixtures/failures/03-fixture-schema-drift.md +16 -16
  92. package/docs/templates/frontend-eval/fixtures/failures/04-missing-loading-empty-error-state.md +16 -16
  93. package/docs/templates/frontend-eval/fixtures/failures/05-forbidden-write-writeset-expansion.md +16 -16
  94. package/docs/templates/frontend-eval/fixtures/failures/06-unapproved-dependency-add.md +16 -16
  95. package/docs/templates/frontend-eval/fixtures/failures/07-mock-production-on.md +21 -21
  96. package/docs/templates/frontend-eval/fixtures/functional/01-simple-component-style.md +29 -29
  97. package/docs/templates/frontend-eval/fixtures/functional/02-form-validation.md +28 -28
  98. package/docs/templates/frontend-eval/fixtures/functional/03-list-detail-page.md +28 -28
  99. package/docs/templates/frontend-eval/fixtures/functional/04-api-mock.md +29 -29
  100. package/docs/templates/frontend-eval/fixtures/functional/05-permission-auth-gated-ui.md +27 -27
  101. package/docs/templates/frontend-eval/fixtures/functional/06-ssr-server-client-boundary.md +28 -28
  102. package/docs/templates/frontend-eval/fixtures/functional/07-shared-public-component-api.md +28 -28
  103. package/docs/templates/frontend-eval/fixtures/functional/08-pure-local-no-remote.md +27 -27
  104. package/docs/templates/frontend-eval/metrics.md +138 -138
  105. package/docs/templates/frontend-eval/smoke-targets.md +53 -53
  106. package/docs/templates/frontend-implementation-contract.schema.json +27 -27
  107. package/docs/templates/frontend-task-constraints.md +35 -35
  108. package/docs/templates/frontend-task-requirement.md +70 -70
  109. package/docs/templates/frontend-test-dag.generate-cases.prompt.md +5 -5
  110. package/docs/templates/frontend-test-dag.json +23 -23
  111. package/docs/templates/frontend-test-dag.retrieve-context.prompt.md +3 -3
  112. package/docs/templates/frontend-test-dag.retrospect.prompt.md +3 -3
  113. package/docs/templates/frontend-test-dag.review-cases.prompt.md +3 -3
  114. package/docs/templates/frontend-test-dag.review-execution.prompt.md +3 -3
  115. package/docs/templates/harness.schema.json +221 -221
  116. package/docs/templates/hybrid-dag.json +188 -188
  117. package/docs/templates/init-evolution-review.md +35 -35
  118. package/docs/templates/interactive-ui-round2-experiment.md +66 -66
  119. package/docs/templates/knowledge-graph-bootstrap-dag.json +118 -118
  120. package/docs/templates/knowledge-sync-dag.json +178 -178
  121. package/docs/templates/knowledge-sync-draft.schema.json +71 -71
  122. package/docs/templates/product-line/AGENTS.md +8 -8
  123. package/docs/templates/product-line/README.md +9 -9
  124. package/docs/templates/product-line/acceptance.yaml +14 -14
  125. package/docs/templates/product-line/closeout.yaml +9 -9
  126. package/docs/templates/product-line/design.md +13 -13
  127. package/docs/templates/product-line/links.md +10 -10
  128. package/docs/templates/product-line/requirement.md +17 -17
  129. package/docs/templates/product-line/task-graph.yaml +15 -15
  130. package/docs/templates/product-line/task.yaml +64 -64
  131. package/docs/templates/product-line/test-plan.md +7 -7
  132. package/docs/templates/production-readiness-checklist.md +57 -57
  133. package/docs/templates/progress-log.md +17 -17
  134. package/docs/templates/project-start-checklist.md +9 -9
  135. package/docs/templates/qa-report.md +48 -48
  136. package/docs/templates/sprint-contract.md +29 -29
  137. package/docs/templates/worker-dogfood-evidence.md +80 -80
  138. package/docs/templates/worker-dogfood-setup.md +68 -68
  139. package/examples/decision-gate-agent-dag.json +173 -173
  140. package/examples/example-dag.json +46 -46
  141. package/examples/hybrid-loop-agent-dag.json +188 -188
  142. package/harness.json +66 -66
  143. package/package.json +78 -52
  144. package/scripts/kb-bootstrap-init-skeleton.sh +240 -240
  145. package/scripts/kb-graph-incremental-prepare.mjs +386 -386
  146. package/scripts/kb-graph-materialize.mjs +105 -105
  147. package/scripts/kb-graph-promote.mjs +164 -164
  148. package/scripts/kb-query.mjs +554 -554
  149. package/skills/agent-worker/SKILL.md +39 -39
  150. package/skills/agent-worker/references/agent-worker-operator.md +60 -60
  151. package/skills/ai-engineering-context/SKILL.md +48 -48
  152. package/skills/analyze-product-dependencies/SKILL.md +67 -67
  153. package/skills/analyze-product-dependencies/agents/openai.yaml +4 -4
  154. package/skills/analyze-product-dependencies/references/api-documentation-schema.md +30 -30
  155. package/skills/analyze-product-dependencies/references/dependency-analysis-schema.md +28 -28
  156. package/skills/analyze-product-dependencies/references/example.md +76 -76
  157. package/skills/analyze-product-dependencies/references/forward-test-cases.md +35 -35
  158. package/skills/analyze-product-dependencies/references/input-contract.md +11 -11
  159. package/skills/analyze-product-dependencies/references/scouting-rules.md +61 -61
  160. package/skills/analyze-product-dependencies/scripts/test-validators.mjs +267 -267
  161. package/skills/analyze-product-dependencies/scripts/validate-api-documentation.mjs +101 -101
  162. package/skills/analyze-product-dependencies/scripts/validate-dependency-analysis.mjs +142 -142
  163. package/skills/analyze-product-dependencies/scripts/validate-product-requirement-input.mjs +76 -76
  164. package/skills/analyze-product-dependencies/scripts/validation-helpers.mjs +146 -146
  165. package/skills/analyze-product-requirements/SKILL.md +90 -90
  166. package/skills/analyze-product-requirements/agents/openai.yaml +4 -4
  167. package/skills/analyze-product-requirements/references/acceptance-criteria.md +91 -91
  168. package/skills/analyze-product-requirements/references/clarification-and-knowledge.md +56 -56
  169. package/skills/analyze-product-requirements/references/example.md +86 -86
  170. package/skills/analyze-product-requirements/references/forward-test-cases.md +66 -66
  171. package/skills/analyze-product-requirements/references/product-analysis-schema.md +32 -32
  172. package/skills/analyze-product-requirements/references/product-requirement-schema.md +33 -33
  173. package/skills/analyze-product-requirements/references/requirement-clarification-schema.md +35 -35
  174. package/skills/analyze-product-requirements/scripts/test-validators.mjs +193 -193
  175. package/skills/analyze-product-requirements/scripts/validate-product-analysis.mjs +69 -69
  176. package/skills/analyze-product-requirements/scripts/validate-product-requirement.mjs +97 -97
  177. package/skills/analyze-product-requirements/scripts/validate-requirement-clarification.mjs +98 -98
  178. package/skills/analyze-product-requirements/scripts/validation-helpers.mjs +156 -156
  179. package/skills/browser-tools/SKILL.md +196 -196
  180. package/skills/browser-tools/browser-content.js +103 -103
  181. package/skills/browser-tools/browser-cookies.js +35 -35
  182. package/skills/browser-tools/browser-eval.js +53 -53
  183. package/skills/browser-tools/browser-hn-scraper.js +108 -108
  184. package/skills/browser-tools/browser-nav.js +44 -44
  185. package/skills/browser-tools/browser-pick.js +162 -162
  186. package/skills/browser-tools/browser-screenshot.js +34 -34
  187. package/skills/browser-tools/browser-start.js +86 -86
  188. package/skills/browser-tools/package-lock.json +2556 -2556
  189. package/skills/browser-tools/package.json +19 -19
  190. package/skills/code-review-core/SKILL.md +20 -20
  191. package/skills/codebase-scout/SKILL.md +19 -19
  192. package/skills/frontend-design-review/SKILL.md +66 -66
  193. package/skills/frontend-design-review/references/review-checklist.md +40 -58
  194. package/skills/frontend-implementation/SKILL.md +49 -49
  195. package/skills/frontend-implementation/references/code-standards.md +32 -32
  196. package/skills/frontend-implementation/references/design-spec.md +46 -46
  197. package/skills/frontend-implementation/references/node-contracts.md +27 -27
  198. package/skills/frontend-review/SKILL.md +61 -59
  199. package/skills/frontend-review/references/review-findings.md +48 -47
  200. package/skills/frontend-verification/SKILL.md +55 -53
  201. package/skills/frontend-verification/references/verification-checklist.md +59 -68
  202. package/skills/grill-me/SKILL.md +10 -10
  203. package/skills/grill-with-docs/SKILL.md +88 -88
  204. package/skills/grill-with-docs/adr-format.md +47 -47
  205. package/skills/grill-with-docs/context-format.md +60 -60
  206. package/skills/init-capability-evolution/SKILL.md +70 -70
  207. package/skills/loop-agent/SKILL.md +151 -151
  208. package/skills/loop-agent/references/README.md +67 -67
  209. package/skills/loop-agent/references/command-reference.md +527 -527
  210. package/skills/loop-agent/references/docs-converge.md +126 -126
  211. package/skills/loop-agent/references/harness-policy.md +263 -263
  212. package/skills/loop-agent/references/hybrid-dag.md +243 -243
  213. package/skills/loop-agent/references/learned/README.md +21 -21
  214. package/skills/loop-agent/references/long-running-loop.md +57 -57
  215. package/skills/loop-agent/references/model-routing.md +36 -36
  216. package/skills/loop-agent/references/multi-worktree.md +54 -54
  217. package/skills/loop-agent/references/one-shot-runs.md +85 -85
  218. package/skills/loop-agent/references/orchestrator-and-interventions.md +169 -169
  219. package/skills/loop-agent/references/pi-prompt.md +23 -23
  220. package/skills/loop-agent/references/pi-subagent-assisted-mode.md +84 -84
  221. package/skills/loop-agent/references/post-implementation-and-patterns.md +44 -44
  222. package/skills/loop-agent/references/task-workflow.md +89 -89
  223. package/skills/loop-agent/references/verification-and-failure-handling.md +141 -141
  224. package/skills/playwright-cli/SKILL.md +420 -420
  225. package/skills/playwright-cli/references/element-attributes.md +23 -23
  226. package/skills/playwright-cli/references/playwright-tests.md +39 -39
  227. package/skills/playwright-cli/references/request-mocking.md +87 -87
  228. package/skills/playwright-cli/references/running-code.md +241 -241
  229. package/skills/playwright-cli/references/session-management.md +225 -225
  230. package/skills/playwright-cli/references/storage-state.md +275 -275
  231. package/skills/playwright-cli/references/test-generation.md +433 -433
  232. package/skills/playwright-cli/references/tracing.md +139 -139
  233. package/skills/playwright-cli/references/video-recording.md +143 -143
  234. package/skills/playwright-cli-case-generator/SKILL.md +74 -74
  235. package/skills/requesting-code-review/SKILL.md +101 -101
  236. package/skills/requesting-code-review/code-reviewer.md +168 -168
  237. package/skills/systematic-debugging/CREATION-LOG.md +119 -119
  238. package/skills/systematic-debugging/SKILL.md +296 -296
  239. package/skills/systematic-debugging/condition-based-waiting-example.ts +158 -158
  240. package/skills/systematic-debugging/condition-based-waiting.md +115 -115
  241. package/skills/systematic-debugging/defense-in-depth.md +122 -122
  242. package/skills/systematic-debugging/find-polluter.sh +63 -63
  243. package/skills/systematic-debugging/root-cause-tracing.md +169 -169
  244. package/skills/systematic-debugging/test-academic.md +14 -14
  245. package/skills/systematic-debugging/test-pressure-1.md +58 -58
  246. package/skills/systematic-debugging/test-pressure-2.md +68 -68
  247. package/skills/systematic-debugging/test-pressure-3.md +69 -69
  248. package/skills/test-driven-development/SKILL.md +20 -20
  249. package/skills/using-git-worktrees/SKILL.md +215 -215
  250. package/skills/verification-before-completion/SKILL.md +154 -154
  251. package/skills/webapp-testing/SKILL.md +19 -19
  252. package/docs/agent-dag-recovery-playbook.md +0 -195
  253. package/docs/agent-dag-runner.md +0 -67
  254. package/docs/cursor-prompt-sidecar.md +0 -36
  255. package/docs/decisions/README.md +0 -18
  256. package/docs/design/README.md +0 -167
  257. package/docs/development-principles.md +0 -73
  258. package/docs/exec-plans/README.md +0 -6
  259. package/docs/exec-plans/active/README.md +0 -12
  260. package/docs/exec-plans/completed/README.md +0 -107
  261. package/docs/feature-workflow.md +0 -414
  262. package/docs/loop-agent-harness.md +0 -142
  263. package/docs/production-readiness.md +0 -96
  264. package/docs/progress/README.md +0 -80
  265. package/docs/reports/README.md +0 -159
  266. package/docs/verification-matrix.md +0 -70
  267. package/scripts/check-product-line-docs.sh +0 -29
  268. package/scripts/check-task-pool-root.sh +0 -32
  269. package/scripts/kb-graph-incremental-prepare.sh +0 -5
  270. package/scripts/kb-graph-materialize.sh +0 -4
  271. package/scripts/kb-graph-promote.sh +0 -4
  272. package/scripts/kb-query.sh +0 -5
@@ -1,27 +1,27 @@
1
- # Frontend Node Contracts
2
-
3
- Pre-write nodes are read-only. Preserve IDs, labels, commands, language, required headings.
4
-
5
- ## Core nodes
6
-
7
- - **`frontend-contract-pi`**: `Scope`, `Non-goals`, `Acceptance Criteria`, `UI States`, `Target Runtime Environment`, `Risks`, `Verification Expectations`. No guessed requirements.
8
- - **`frontend-scout-pi`**: routes, components, tokens, data/API/Mock, scripts, tests, assets. Fact vs inference vs gap. Knowledge base first; else search+read `<repoRoot>/openSpec/**` before repo fallback. Output stack, routes, components, styling, conventions, state/data, test entry points, reuse, risks.
9
- - **`frontend-mock-assess-pi` + gate**: first non-empty line
10
- `MOCK_STRATEGY: native|browser-intercept|request-adapter|not-needed|blocked`
11
- Prefer native Mock; browser intercept only with existing e2e; request-adapter only for reversible local preview. `not-needed` needs positive no-remote/stable-backend evidence; invalid when `frontendMock.policy=required`. `blocked` for missing/conflicting contracts, unsafe paths/deps, unread specs, production-default-on, unverifiable entrypoints. Output Mock Decision, API/spec/service evidence, backend readiness, selection evidence, endpoint/fixture matrix, activation, targets, production safety, verification plan, real-integration gap, blocking issues. Never invent fields, store secrets, comment real requests, import test mocks into production, or treat Mock as real integration. Gate uses `first-non-empty` only; never authorizes writes. Unsafe required contracts → no writer.
12
- - **`frontend-plan-pi` + design loop**: AC → steps, in-bound files, UI states, reuse, deps, activation/rollback, frozen verify entrypoints, real-integration gap. First gate: `VERDICT: pass|request-revision`. Pass may emit `PASS_NO_REVISION_NEEDED`; else full corrected plan without invented evidence. Final review rechecks plan/findings/revision/assessment/Mock safety. Only final `VERDICT: pass` authorizes writes; failure → replan/rerun (not dev-fix).
13
- - **`frontend-implement-pi`**: sole exclusive writer. Stay in `writeSet`; real requests default-on; Mock reversible, dev/test-only, production-off. Atomic handler/intercept/adapter with consumer+tests. Stop on forbidden paths or guesses. Output changed files, behavior, UI states, styling notes, verification attempted, residual risks. Optional mock-verify when frozen; static+behavior always; behavior must prove page consumption.
14
-
15
- ## Contract / trace / stages (M1–M2)
16
-
17
- - Contract shell: `jsonArtifactGate` → `contracts/frontend-implementation-contract.json` from revision (one fenced JSON or pure JSON); schemaId `frontend-implementation-contract-v1`. Final design + implement depend on it; `MOCK_STRATEGY: blocked` not implementable.
18
- - Trace shell: `frontend-verification-trace-gate` binds contract `verificationTargets` to static/behavior records (`commandLabels`, file exists, optional symbol). Assess `MOCK_STRATEGY:` must match `mockApi.strategy` when present. Writes `contracts/frontend-verification-trace.json`. Browser/visual always `not-run`.
19
- - Implement stages: (1) contract confirm (2) tests sync (3) component/UI (4) API/Mock (5) frozen checks (6) diff cleanup. Summary: Contract Ref, Changed Files, Requirements, UI States, Tests, Verification Attempts, Deviations, Residual Risks.
20
-
21
- ## Repair (M3)
22
-
23
- static/behavior/trace may `nonZeroExitPolicy: record`. Assess → `contracts/frontend-repair-assessment.json`. Repair-contract fail-closed on non-repairable (contract/path/dependency/credential/deploy/spec-unclear) or writeSet expansion. `frontend-repair-pi`: same writeSet as implement; no re-spec; max 1 attempt. Reverify/retrace fail policy; review/closeout use post-repair evidence.
24
-
25
- ## Risk & capability (M4–M6)
26
-
27
- Deterministic risk (no model); high-risk beats small; supervised never small. Small may drop first design gate + plan-revision; contract shell retargets to `frontend-plan-pi`. Capability seed injects adapters; openSpec/task sources outrank. A11y: static/component tools only when present; Browser a11y always not-run.
1
+ # Frontend Node Contracts
2
+
3
+ Pre-write nodes are read-only. Preserve IDs, labels, commands, language, required headings.
4
+
5
+ ## Core nodes
6
+
7
+ - **`frontend-contract-pi`**: `Scope`, `Non-goals`, `Acceptance Criteria`, `UI States`, `Target Runtime Environment`, `Risks`, `Verification Expectations`. No guessed requirements.
8
+ - **`frontend-scout-pi`**: routes, components, tokens, data/API/Mock, scripts, tests, assets. Fact vs inference vs gap. Knowledge base first; else search+read `<repoRoot>/openSpec/**` before repo fallback. Output stack, routes, components, styling, conventions, state/data, test entry points, reuse, risks.
9
+ - **`frontend-mock-assess-pi` + gate**: first non-empty line
10
+ `MOCK_STRATEGY: native|browser-intercept|request-adapter|not-needed|blocked`
11
+ Prefer native Mock; browser intercept only with existing e2e; request-adapter only for reversible local preview. Default `auto` may select `not-needed` when contract/scout evidence confirms no project Mock capability, without adding Mock files/deps, while keeping real requests default and recording the Real Integration Gap. Other `not-needed` cases need positive no-remote/stable-backend evidence; invalid when `frontendMock.policy=required`. `blocked` for missing/conflicting contracts, unsafe paths/deps, unread specs, production-default-on, unverifiable entrypoints. Output Mock Decision, API/spec/service evidence, backend readiness, selection evidence, endpoint/fixture matrix, activation, targets, production safety, verification plan, real-integration gap, blocking issues. Never invent fields, store secrets, comment real requests, import test mocks into production, or treat Mock as real integration. Gate uses `first-non-empty` only; never authorizes writes. Unsafe required contracts → no writer.
12
+ - **`frontend-plan-pi` + design loop**: AC → steps, in-bound files, UI states, reuse, deps, activation/rollback, frozen verify entrypoints, real-integration gap. First gate: `VERDICT: pass|request-revision`. Pass may emit `PASS_NO_REVISION_NEEDED`; else full corrected plan without invented evidence. Final review rechecks plan/findings/revision/assessment/Mock safety. Only final `VERDICT: pass` authorizes writes; failure → replan/rerun (not dev-fix).
13
+ - **`frontend-implement-pi`**: sole exclusive writer. Stay in `writeSet`; real requests default-on; Mock reversible, dev/test-only, production-off. Atomic handler/intercept/adapter with consumer+tests. Stop on forbidden paths or guesses. Output changed files, behavior, UI states, styling notes, verification attempted, residual risks. Optional mock-verify when frozen; static+behavior always; behavior must prove page consumption. Skipped-Mock `not-needed` keeps real integration pending unless the real backend path has fresh evidence.
14
+
15
+ ## Contract / trace / stages (M1–M2)
16
+
17
+ - Contract shell: `jsonArtifactGate` → `contracts/frontend-implementation-contract.json` from revision (one fenced JSON or pure JSON); schemaId `frontend-implementation-contract-v1`. Final design + implement depend on it; `MOCK_STRATEGY: blocked` not implementable.
18
+ - Trace shell: `frontend-verification-trace-gate` binds contract `verificationTargets` to static/behavior records (`commandLabels`, file exists, optional symbol). Assess `MOCK_STRATEGY:` must match `mockApi.strategy` when present. Writes `contracts/frontend-verification-trace.json`. Browser/visual always `not-run`.
19
+ - Implement stages: (1) contract confirm (2) tests sync (3) component/UI (4) API/Mock (5) frozen checks (6) diff cleanup. Summary: Contract Ref, Changed Files, Requirements, UI States, Tests, Verification Attempts, Deviations, Residual Risks.
20
+
21
+ ## Repair (M3)
22
+
23
+ static/behavior/trace may `nonZeroExitPolicy: record`. Assess → `contracts/frontend-repair-assessment.json`. Repair-contract fail-closed on non-repairable (contract/path/dependency/credential/deploy/spec-unclear) or writeSet expansion. `frontend-repair-pi`: same writeSet as implement; no re-spec; max 1 attempt. Reverify/retrace fail policy; review/closeout use post-repair evidence.
24
+
25
+ ## Risk & capability (M4–M6)
26
+
27
+ Deterministic risk (no model); high-risk beats small; supervised never small. Small may drop first design gate + plan-revision; contract shell retargets to `frontend-plan-pi`. Capability seed injects adapters; openSpec/task sources outrank. A11y: static/component tools only when present; Browser a11y always not-run.
@@ -1,59 +1,61 @@
1
- ---
2
- name: frontend-review
3
- description: Use to review completed frontend code and verification before closeout.
4
- references:
5
- - path: references/review-findings.md
6
- required: true
7
- maxChars: 2800
8
- ---
9
-
10
- # Frontend Review
11
-
12
- Use for `frontend-review-pi`; read the findings guide first. Required inputs are
13
- original task/reference material, contract/constraints, Mock assessment, original and
14
- revised/confirmed plan, final design verdict, implementation summary, actual diff,
15
- and static/behavior/optional Mock shell evidence. Missing actual diff or required
16
- evidence forces revision; never infer it from a summary.
17
-
18
- ## Verdict Contract
19
-
20
- The first non-empty line must be exactly `VERDICT: pass` or
21
- `VERDICT: request-revision`. Any Critical/Important finding, failed or missing
22
- required check, forbidden write, or unmet acceptance criterion forces revision.
23
-
24
- ## Review Scope
25
-
26
- - Compare intent, contract, plan, diff, and evidence; report altered requirements.
27
- - Inspect every changed file against allowed, forbidden, and approved write scope.
28
- - Map criteria to behavior, applicable UI states, tests, and shell evidence.
29
- - Review state/data flow, validation, async/error behavior, components/design, responsive behavior, accessibility, dependencies, maintenance, and regression risk when applicable.
30
- - Inspect static, behavior, and available Mock-specific artifacts directly. For Mock
31
- strategies, compare the endpoint matrix, handler/fixture/adapter and consumer diff;
32
- require the real request as default, contract-aligned fixtures, production isolation,
33
- and no false real-integration claim. `not-needed` needs applicable real/no-remote evidence.
34
- - Component/design claims require traceable knowledge-base evidence or, after connection/query failure or no match, relevant `<repoRoot>/openSpec/**` evidence. The connector format is TODO; never claim a query or fallback search without evidence. Execute explicit `grep`/`find` to locate spec files and `read` to load them before referencing their rules. Only successful `read` tool calls are observable as "已读取规范文件" in the spec-evidence inspector.
35
- - Treat shell exit status as authoritative. Do not edit files.
36
-
37
- ## Evidence And Output
38
-
39
- Findings cite a tight file location, exact command/result, or named DAG artifact.
40
- Separate confirmed defects, missing evidence, and residual risks.
41
-
42
- ```markdown
43
- VERDICT: request-revision
44
-
45
- ## Findings
46
- - [Important] `path:line` — issue, impact, and required correction.
47
-
48
- ## Verification Assessment
49
- - ...
50
-
51
- ## UX Assessment
52
- - ...
53
-
54
- ## Residual Risks
55
- - ...
56
- ```
57
-
58
- A pass requires no Critical/Important findings and all required shell checks passed.
59
- Still report knowledge-source status and optional browser/manual gaps.
1
+ ---
2
+ name: frontend-review
3
+ description: Use to review completed frontend code and verification before closeout.
4
+ references:
5
+ - path: references/review-findings.md
6
+ required: true
7
+ maxChars: 2800
8
+ ---
9
+
10
+ # Frontend Review
11
+
12
+ Use for `frontend-review-pi`; read the findings guide first. Required inputs are
13
+ original task/reference material, contract/constraints, Mock assessment, original and
14
+ revised/confirmed plan, final design verdict, implementation summary, actual diff,
15
+ and static/behavior/optional Mock shell evidence. Missing actual diff or required
16
+ evidence forces revision; never infer it from a summary.
17
+
18
+ ## Verdict Contract
19
+
20
+ The first non-empty line must be exactly `VERDICT: pass` or
21
+ `VERDICT: request-revision`. Any Critical/Important finding, failed or missing
22
+ required check, forbidden write, or unmet acceptance criterion forces revision.
23
+
24
+ ## Review Scope
25
+
26
+ - Compare intent, contract, plan, diff, and evidence; report altered requirements.
27
+ - Inspect every changed file against allowed, forbidden, and approved write scope.
28
+ - Map criteria to behavior, applicable UI states, tests, and shell evidence.
29
+ - Review state/data flow, validation, async/error behavior, components/design, responsive behavior, accessibility, dependencies, maintenance, and regression risk when applicable.
30
+ - Inspect static, behavior, and available Mock-specific artifacts directly. For Mock
31
+ strategies, compare the endpoint matrix, handler/fixture/adapter and consumer diff;
32
+ require the real request as default, contract-aligned fixtures, production isolation,
33
+ and no false real-integration claim. `not-needed` needs applicable real/no-remote
34
+ evidence, or an explicit default-auto skipped-Mock rationale with the Real
35
+ Integration Gap preserved when no project Mock capability is confirmed.
36
+ - Component/design claims require traceable knowledge-base evidence or, after connection/query failure or no match, relevant `<repoRoot>/openSpec/**` evidence. The connector format is TODO; never claim a query or fallback search without evidence. Execute explicit `grep`/`find` to locate spec files and `read` to load them before referencing their rules. Only successful `read` tool calls are observable as "已读取规范文件" in the spec-evidence inspector.
37
+ - Treat shell exit status as authoritative. Do not edit files.
38
+
39
+ ## Evidence And Output
40
+
41
+ Findings cite a tight file location, exact command/result, or named DAG artifact.
42
+ Separate confirmed defects, missing evidence, and residual risks.
43
+
44
+ ```markdown
45
+ VERDICT: request-revision
46
+
47
+ ## Findings
48
+ - [Important] `path:line` — issue, impact, and required correction.
49
+
50
+ ## Verification Assessment
51
+ - ...
52
+
53
+ ## UX Assessment
54
+ - ...
55
+
56
+ ## Residual Risks
57
+ - ...
58
+ ```
59
+
60
+ A pass requires no Critical/Important findings and all required shell checks passed.
61
+ Still report knowledge-source status and optional browser/manual gaps.
@@ -1,47 +1,48 @@
1
- # Frontend Review Findings Guide
2
-
3
- ## Severity
4
-
5
- - **Critical**: blocks primary flow, corrupts data, violates security/privacy, writes forbidden paths, or bypasses required verification.
6
- - **Important**: acceptance/state/validation gap, material convention drift, missing behavior tests, unauthorized dependency, unsafe mock activation/import, mock-contract drift, misleading real-integration claim, or failed/missing required verification.
7
- - **Minor**: non-blocking maintainability, copy, layout, or cleanup issue.
8
-
9
- ## Evidence
10
-
11
- - Cite tight file locations, exact commands/results, or named DAG artifacts.
12
- - Never invent evidence; name the missing check. An implementation summary is not the actual diff.
13
- - Failed required static/behavior verification is at least Important unless proven unrelated.
14
- - A knowledge-base claim records connector/query, source ID/version, and retrieval time. If absent, failed, or unmatched, review evidence must show `<repoRoot>/openSpec/**` search terms and matched paths/headings; label `openSpec fallback`, `repository fallback`, or `unavailable` accurately.
15
-
16
- ## Review Sequence
17
-
18
- 1. Establish changed-file inventory and write boundaries.
19
- 2. Compare original requirement with derived contract/constraints.
20
- 3. Map each criterion to code, states, tests, and evidence.
21
- 4. Inspect interactions, state/data/API behavior, failure paths, and regression risk.
22
- 5. Check mock selection, contract-to-fixture mapping, activation/default path,
23
- handler/fixture/adapter and consumer diff, optional Mock-specific verification,
24
- production imports, evidence scope, and the documented real-integration gap.
25
- 6. Check component/design evidence, responsive/accessibility behavior, dependencies, and maintenance fit when applicable.
26
- 7. Classify findings and derive the verdict mechanically.
27
-
28
- Skipping the required `openSpec/` search after knowledge-base failure is Important
29
- when component/design compliance affects acceptance or implementation choices.
30
-
31
- Use one issue per finding:
32
-
33
- ```text
34
- - [Critical|Important|Minor] path:line — Problem; impact; required correction; evidence.
35
- ```
36
-
37
- Avoid vague advice. When no source location exists, cite the command or artifact.
38
-
39
- ## Pass Rules
40
-
41
- - No Critical or Important findings remain.
42
- - Required static and behavior nodes ran and passed.
43
- - Changed files are authorized.
44
- - Criteria and applicable states have implementation and evidence.
45
- - Required Mock-backed behavior passed; any generated Mock-specific verification also
46
- passed; Mock is not enabled by default in production.
47
- - Optional unavailable knowledge-base, browser, visual, or manual checks remain explicit risks.
1
+ # Frontend Review Findings Guide
2
+
3
+ ## Severity
4
+
5
+ - **Critical**: blocks primary flow, corrupts data, violates security/privacy, writes forbidden paths, or bypasses required verification.
6
+ - **Important**: acceptance/state/validation gap, material convention drift, missing behavior tests, unauthorized dependency, unsafe mock activation/import, mock-contract drift, misleading real-integration claim, or failed/missing required verification.
7
+ - **Minor**: non-blocking maintainability, copy, layout, or cleanup issue.
8
+
9
+ ## Evidence
10
+
11
+ - Cite tight file locations, exact commands/results, or named DAG artifacts.
12
+ - Never invent evidence; name the missing check. An implementation summary is not the actual diff.
13
+ - Failed required static/behavior verification is at least Important unless proven unrelated.
14
+ - A knowledge-base claim records connector/query, source ID/version, and retrieval time. If absent, failed, or unmatched, review evidence must show `<repoRoot>/openSpec/**` search terms and matched paths/headings; label `openSpec fallback`, `repository fallback`, or `unavailable` accurately.
15
+
16
+ ## Review Sequence
17
+
18
+ 1. Establish changed-file inventory and write boundaries.
19
+ 2. Compare original requirement with derived contract/constraints.
20
+ 3. Map each criterion to code, states, tests, and evidence.
21
+ 4. Inspect interactions, state/data/API behavior, failure paths, and regression risk.
22
+ 5. Check mock selection, contract-to-fixture mapping, activation/default path,
23
+ handler/fixture/adapter and consumer diff, optional Mock-specific verification,
24
+ production imports, evidence scope, and the documented real-integration gap.
25
+ 6. Check component/design evidence, responsive/accessibility behavior, dependencies, and maintenance fit when applicable.
26
+ 7. Classify findings and derive the verdict mechanically.
27
+
28
+ Skipping the required `openSpec/` search after knowledge-base failure is Important
29
+ when component/design compliance affects acceptance or implementation choices.
30
+
31
+ Use one issue per finding:
32
+
33
+ ```text
34
+ - [Critical|Important|Minor] path:line — Problem; impact; required correction; evidence.
35
+ ```
36
+
37
+ Avoid vague advice. When no source location exists, cite the command or artifact.
38
+
39
+ ## Pass Rules
40
+
41
+ - No Critical or Important findings remain.
42
+ - Required static and behavior nodes ran and passed.
43
+ - Changed files are authorized.
44
+ - Criteria and applicable states have implementation and evidence.
45
+ - Required Mock-backed behavior passed; any generated Mock-specific verification also
46
+ passed; Mock is not enabled by default in production. For default-auto skipped
47
+ Mock, the real request remains default and the Real Integration Gap is preserved.
48
+ - Optional unavailable knowledge-base, browser, visual, or manual checks remain explicit risks.
@@ -1,53 +1,55 @@
1
- ---
2
- name: frontend-verification
3
- description: Use to assess frontend evidence and produce closeout.
4
- references:
5
- - path: references/verification-checklist.md
6
- required: true
7
- maxChars: 3000
8
- ---
9
-
10
- # Frontend Verification
11
-
12
- Use for `frontend-closeout-pi`. Shell nodes execute commands; this read-only skill
13
- assesses their evidence. Read the checklist first.
14
-
15
- Inputs: criteria, change inventory, shell commands/status/artifacts, review
16
- verdict/findings, and required browser, visual, manual, or knowledge evidence.
17
-
18
- ## Evidence Rules
19
-
20
- - Static evidence covers type/lint/build/schema; behavior evidence must exercise the flow.
21
- - Shell exit status is authoritative. Classify as `passed`, `failed`, `not-run`,
22
- `blocked`, or `unavailable`; only passed satisfies a required check.
23
- - Never use static success as behavior proof, or tests as visual/browser proof they did not exercise.
24
- - Mock-backed behavior proves frontend rendering and state transitions only. It never
25
- proves backend readiness, transport compatibility, or real API integration.
26
- - Unavailable commands remain gaps.
27
- - Resolve design evidence via knowledge base, then `<repoRoot>/openSpec/**` after
28
- failure/no match. Its connector format remains TODO; never invent it. An applied
29
- `openSpec fallback` is available project evidence.
30
- - Separate Mock service/handler checks from page consumption and record the
31
- dev/test-only boundary; handler tests alone do not prove page use.
32
-
33
- ## Method And Output
34
-
35
- Map required checks to fresh evidence, classify gaps, confirm review pass, and state
36
- only proven changes.
37
-
38
- Return Markdown headings:
39
-
40
- - `Changes`: changed behavior and areas.
41
- - `Mock Decision`, `Mock Files`, `Mock Verification`, `Production Boundary`: status is `passed`, `failed`, `not-required`, `blocked`, or `unavailable`.
42
- - `Verification Evidence`: table of check, command/source, status, and artifact/result.
43
- - `Review Result`: exact review verdict and findings.
44
- - `Known Risks`: missing optional checks and environment caveats.
45
- - `Follow-up`: concrete work or `None`.
46
-
47
- When the backend remains unavailable but required mock-backed checks pass, state
48
- `Frontend status: mock-validated` and `Real integration: pending`. Use a completed
49
- `<task-id>-real-api-integration-verify` task before changing the latter to complete;
50
- the follow-up is explicit, not auto-created or auto-executed.
51
-
52
- Do not edit files. Do not claim complete when review is not pass or a required check
53
- is failed, not-run, blocked, unavailable, stale, or contradicted.
1
+ ---
2
+ name: frontend-verification
3
+ description: Use to assess frontend evidence and produce closeout.
4
+ references:
5
+ - path: references/verification-checklist.md
6
+ required: true
7
+ maxChars: 3000
8
+ ---
9
+
10
+ # Frontend Verification
11
+
12
+ Use for `frontend-closeout-pi`. Shell nodes execute commands; this read-only skill
13
+ assesses their evidence. Read the checklist first.
14
+
15
+ Inputs: criteria, change inventory, shell commands/status/artifacts, review
16
+ verdict/findings, and required browser, visual, manual, or knowledge evidence.
17
+
18
+ ## Evidence Rules
19
+
20
+ - Static evidence covers type/lint/build/schema; behavior evidence must exercise the flow.
21
+ - Shell exit status is authoritative. Classify as `passed`, `failed`, `not-run`,
22
+ `blocked`, or `unavailable`; only passed satisfies a required check.
23
+ - Never use static success as behavior proof, or tests as visual/browser proof they did not exercise.
24
+ - Mock-backed behavior proves frontend rendering and state transitions only. It never
25
+ proves backend readiness, transport compatibility, or real API integration.
26
+ - Unavailable commands remain gaps.
27
+ - Resolve design evidence via knowledge base, then `<repoRoot>/openSpec/**` after
28
+ failure/no match. Its connector format remains TODO; never invent it. An applied
29
+ `openSpec fallback` is available project evidence.
30
+ - Separate Mock service/handler checks from page consumption and record the
31
+ dev/test-only boundary; handler tests alone do not prove page use.
32
+
33
+ ## Method And Output
34
+
35
+ Map required checks to fresh evidence, classify gaps, confirm review pass, and state
36
+ only proven changes.
37
+
38
+ Return Markdown headings:
39
+
40
+ - `Changes`: changed behavior and areas.
41
+ - `Mock Decision`, `Mock Files`, `Mock Verification`, `Production Boundary`: status is `passed`, `failed`, `not-required`, `blocked`, or `unavailable`.
42
+ - `Verification Evidence`: table of check, command/source, status, and artifact/result.
43
+ - `Review Result`: exact review verdict and findings.
44
+ - `Known Risks`: missing optional checks and environment caveats.
45
+ - `Follow-up`: concrete work or `None`.
46
+
47
+ When the backend remains unavailable but required mock-backed checks pass, state
48
+ `Frontend status: mock-validated` and `Real integration: pending`. When default
49
+ `auto` skipped Mock and no real API evidence passed, state
50
+ `Frontend status: locally-validated` and `Real integration: pending`. Use a completed
51
+ `<task-id>-real-api-integration-verify` task before changing the latter to complete;
52
+ the follow-up is explicit, not auto-created or auto-executed.
53
+
54
+ Do not edit files. Do not claim complete when review is not pass or a required check
55
+ is failed, not-run, blocked, unavailable, stale, or contradicted.
@@ -1,68 +1,59 @@
1
- # Frontend Verification Checklist
2
-
3
- ## Static And Behavior Evidence
4
-
5
- - Type/compile, lint/format, build, and schema/client checks ran when required.
6
- - Generated output was authorized.
7
- - Tests cover changed logic/flows and regressions.
8
- - Browser/e2e/manual evidence exists when explicitly required.
9
- - Selected-strategy behavior and applicable loading, empty, error, success, disabled,
10
- permission, retry, and boundary states have evidence from fixed DAG entrypoints.
11
- - When generated, Mock-specific verification checks service/handler/schema/fixtures;
12
- behavior evidence separately proves page consumption.
13
- - For Mock strategies, activation is explicit/non-production and a production/default-
14
- real-path build with Mock off confirms the real request remains default.
15
- - `not-needed` has positive readiness/no-remote evidence plus applicable real or
16
- no-remote behavior evidence. Mock-backed evidence remains frontend-only and never
17
- satisfies real API integration.
18
-
19
- ## Design And Component Evidence
20
-
21
- - Claims cite knowledge-base retrieval or `<repoRoot>/openSpec/**` fallback.
22
- - Evidence records query/source/version/time or fallback search terms, paths, headings, and applied rules.
23
- - Relevant `openSpec/` matches become the current-project specification and satisfy source availability; only the knowledge-base connection remains unavailable.
24
- - Missing both sources blocks explicit compliance or an unresolved required design decision.
25
-
26
- ## Status
27
-
28
- - `passed`: fresh successful evidence matches current implementation.
29
- - `failed`: check ran and failed.
30
- - `not-run`: no fresh attempt exists.
31
- - `blocked`: a prerequisite prevented execution.
32
- - `unavailable`: tool, environment, connector, or source was absent.
33
-
34
- Only passed satisfies a required check. Other optional statuses remain disclosed risks.
35
-
36
- ## Closeout Checks
37
-
38
- - List exact commands, source, exit status, and archived output/artifact when available.
39
- - Map every criterion to evidence or a named gap.
40
- - Record review verdict before completion.
41
- - Do not conflate static, behavior, browser/visual/manual, or knowledge-base proof.
42
- - If only mock evidence exists, report `Frontend status: mock-validated` and
43
- `Real integration: pending`, with the actual API verification as follow-up.
44
-
45
- ```markdown
46
- ## Changes
47
- - ...
48
-
49
- ## Verification Evidence
50
- | Check | Command or source | Status | Evidence |
51
- |---|---|---|---|
52
- | ... | ... | passed | ... |
53
-
54
- ## Mock Decision / Mock Files / Mock Verification / Production Boundary
55
- - Status: `passed | failed | not-required | blocked | unavailable`
56
-
57
- ## Review Result
58
- - Verdict: `VERDICT: pass`
59
-
60
- ## Known Risks
61
- - ...
62
-
63
- ## Follow-up
64
- - None.
65
- ```
66
-
67
- If review is not pass or a required check is not passed, describe the task as
68
- incomplete and list concrete follow-up.
1
+ # Frontend Verification Checklist
2
+
3
+ ## Static And Behavior Evidence
4
+
5
+ - Required type/compile, lint/format, build, schema/client, browser/e2e/manual checks ran.
6
+ - Generated output was authorized; tests cover changed logic, flows, and regressions.
7
+ - Fixed DAG entrypoints prove selected-strategy behavior and applicable loading/empty/error/success/disabled/permission/retry/boundary states.
8
+ - Mock-specific checks cover service/handler/schema/fixtures; behavior evidence separately proves page consumption.
9
+ - Mock activation is explicit/non-production; a default-real-path build with Mock off proves the real request remains default.
10
+ - `not-needed` has real/no-remote evidence, or default-auto skipped-Mock rationale with Real Integration Gap preserved when no project Mock capability exists.
11
+ - Mock-backed evidence is frontend-only and never satisfies real API integration.
12
+
13
+ ## Design And Component Evidence
14
+
15
+ - Claims cite knowledge-base retrieval or `<repoRoot>/openSpec/**` fallback.
16
+ - Evidence records query/source/time or fallback search terms, paths, headings, and applied rules.
17
+ - Relevant `openSpec/` matches satisfy source availability; missing both sources blocks explicit compliance or required design decisions.
18
+
19
+ ## Status
20
+
21
+ - `passed`: fresh successful evidence matches current implementation.
22
+ - `failed`: check ran and failed.
23
+ - `not-run`: no fresh attempt exists.
24
+ - `blocked`: a prerequisite prevented execution.
25
+ - `unavailable`: tool, environment, connector, or source was absent.
26
+
27
+ Only passed satisfies a required check. Other optional statuses remain disclosed risks.
28
+
29
+ ## Closeout Checks
30
+
31
+ - List exact commands/source/exit status/artifacts; map every criterion to evidence or a named gap.
32
+ - Record review verdict before completion.
33
+ - Do not conflate static, behavior, browser/visual/manual, or knowledge-base proof.
34
+ - If only Mock evidence exists, report `Frontend status: mock-validated` and `Real integration: pending`, with actual API verification follow-up.
35
+ - If default `auto` skipped Mock and no real API evidence exists, report `Frontend status: locally-validated` and `Real integration: pending`.
36
+
37
+ ```markdown
38
+ ## Changes
39
+ - ...
40
+
41
+ ## Verification Evidence
42
+ | Check | Command or source | Status | Evidence |
43
+ |---|---|---|---|
44
+ | ... | ... | passed | ... |
45
+
46
+ ## Mock Decision / Mock Files / Mock Verification / Production Boundary
47
+ - Status: `passed | failed | not-required | blocked | unavailable`
48
+
49
+ ## Review Result
50
+ - Verdict: `VERDICT: pass`
51
+
52
+ ## Known Risks
53
+ - ...
54
+
55
+ ## Follow-up
56
+ - None.
57
+ ```
58
+
59
+ If review is not pass or a required check is not passed, describe the task as incomplete and list concrete follow-up.
@@ -1,10 +1,10 @@
1
- ---
2
- name: grill-me
3
- description: Interview the user relentlessly about a plan or design until reaching shared understanding, resolving each branch of the decision tree. Use when user wants to stress-test a plan, get grilled on their design, or mentions "grill me".
4
- ---
5
-
6
- Interview me relentlessly about every aspect of this plan until we reach a shared understanding. Walk down each branch of the design tree, resolving dependencies between decisions one-by-one. For each question, provide your recommended answer.
7
-
8
- Ask the questions one at a time.
9
-
10
- If a question can be answered by exploring the codebase, explore the codebase instead.
1
+ ---
2
+ name: grill-me
3
+ description: Interview the user relentlessly about a plan or design until reaching shared understanding, resolving each branch of the decision tree. Use when user wants to stress-test a plan, get grilled on their design, or mentions "grill me".
4
+ ---
5
+
6
+ Interview me relentlessly about every aspect of this plan until we reach a shared understanding. Walk down each branch of the design tree, resolving dependencies between decisions one-by-one. For each question, provide your recommended answer.
7
+
8
+ Ask the questions one at a time.
9
+
10
+ If a question can be answered by exploring the codebase, explore the codebase instead.