@tea-agent/loop-agent 0.13.0 → 0.14.0

This diff represents the content of publicly available package versions that have been released to one of the supported registries. The information contained in this diff is provided for informational purposes only and reflects changes between package versions as they appear in their respective public registries.
Files changed (270) hide show
  1. package/AGENTS.md +157 -157
  2. package/CHANGELOG.md +73 -301
  3. package/README.md +338 -334
  4. package/bin/agent-worker.js +22 -22
  5. package/bin/loop-agent.js +21 -21
  6. package/dist/commands/cursor-prompt.js +6 -6
  7. package/dist/commands/init.js +505 -505
  8. package/dist/commands/loop-benchmark.js +11 -11
  9. package/dist/commands/pi-reuse-benchmark.js +16 -16
  10. package/dist/executors/pi-event-serializer.js +33 -11
  11. package/dist/sidecars/cursor-prompt/executor.js +1 -1
  12. package/dist/task/runtime.js +27 -27
  13. package/dist/worker/observe/spec-evidence.js +19 -10
  14. package/dist/worker/observe/static/api.js +46 -46
  15. package/dist/worker/observe/static/app.js +151 -150
  16. package/dist/worker/observe/static/constants.js +156 -148
  17. package/dist/worker/observe/static/copy.js +67 -67
  18. package/dist/worker/observe/static/dag-helpers.js +201 -172
  19. package/dist/worker/observe/static/dag-layout.d.ts +31 -31
  20. package/dist/worker/observe/static/dag-layout.js +83 -83
  21. package/dist/worker/observe/static/dag-model.js +72 -72
  22. package/dist/worker/observe/static/dom.js +122 -122
  23. package/dist/worker/observe/static/format-pool.d.ts +71 -0
  24. package/dist/worker/observe/static/format-pool.js +134 -67
  25. package/dist/worker/observe/static/format.js +317 -292
  26. package/dist/worker/observe/static/index.html +350 -308
  27. package/dist/worker/observe/static/kpi.js +100 -94
  28. package/dist/worker/observe/static/markdown-render.js +124 -0
  29. package/dist/worker/observe/static/relations.js +133 -133
  30. package/dist/worker/observe/static/router.js +93 -93
  31. package/dist/worker/observe/static/run-processing.js +148 -148
  32. package/dist/worker/observe/static/shell-chrome.js +74 -68
  33. package/dist/worker/observe/static/state.js +273 -267
  34. package/dist/worker/observe/static/styles.css +2504 -1902
  35. package/dist/worker/observe/static/views/batch.js +227 -227
  36. package/dist/worker/observe/static/views/dag-graph.js +172 -172
  37. package/dist/worker/observe/static/views/dag-inspector.js +530 -627
  38. package/dist/worker/observe/static/views/dag.js +371 -371
  39. package/dist/worker/observe/static/views/dashboard.js +86 -100
  40. package/dist/worker/observe/static/views/failures.js +143 -143
  41. package/dist/worker/observe/static/views/feature.js +492 -492
  42. package/dist/worker/observe/static/views/pool.js +708 -350
  43. package/dist/worker/observe/static/views/run.js +453 -453
  44. package/dist/worker/observe/static/views/session-timeline.js +771 -219
  45. package/dist/worker/observe/static/views/shell.js +7 -7
  46. package/dist/worker/observe/static/views/task.js +314 -314
  47. package/dist/worker/observe/static/views/timeline.js +163 -163
  48. package/dist/workflows/dag/canvas-observer.js +275 -275
  49. package/docs/README.md +105 -104
  50. package/docs/agent-dag-recovery-playbook.md +195 -195
  51. package/docs/agent-dag-runner.md +67 -67
  52. package/docs/architecture/README.md +26 -26
  53. package/docs/architecture/dag-execution.md +140 -140
  54. package/docs/architecture/evolution.md +54 -54
  55. package/docs/architecture/facts-and-state.md +71 -71
  56. package/docs/architecture/runtime-boundaries.md +191 -191
  57. package/docs/architecture/system-overview.md +93 -93
  58. package/docs/architecture/worker-and-feature.md +85 -85
  59. package/docs/cursor-prompt-sidecar.md +36 -36
  60. package/docs/decisions/README.md +18 -18
  61. package/docs/design/README.md +167 -167
  62. package/docs/development-principles.md +73 -73
  63. package/docs/exec-plans/README.md +6 -6
  64. package/docs/exec-plans/active/README.md +2 -1
  65. package/docs/exec-plans/completed/README.md +105 -104
  66. package/docs/feature-workflow.md +414 -414
  67. package/docs/harness-methodology-debugging.md +153 -153
  68. package/docs/harness-methodology-tdd.md +130 -130
  69. package/docs/harness-methodology-verification.md +27 -27
  70. package/docs/init-surface.manifest.json +307 -307
  71. package/docs/loop-agent-harness.md +142 -142
  72. package/docs/production-readiness.md +96 -96
  73. package/docs/progress/README.md +59 -58
  74. package/docs/reports/README.md +123 -119
  75. package/docs/skills/README.md +7 -7
  76. package/docs/skills/vetted-skill-registry.md +29 -29
  77. package/docs/templates/adr.md +60 -60
  78. package/docs/templates/agent-dag-authority-surface-audit.prompt.md +94 -94
  79. package/docs/templates/agent-dag-decision-envelope.schema.json +213 -213
  80. package/docs/templates/agent-dag-decision-gate-dogfood-report.md +117 -117
  81. package/docs/templates/agent-dag-decision-gate.prompt.md +246 -246
  82. package/docs/templates/agent-dag-process-supervisor.prompt.md +98 -98
  83. package/docs/templates/agent-dag-report.schema.json +473 -473
  84. package/docs/templates/agent-dag-review-verdict.prompt.md +68 -68
  85. package/docs/templates/agent-dag.base.json +190 -190
  86. package/docs/templates/agent-dag.final-verification.json +185 -185
  87. package/docs/templates/agent-dag.schema.json +411 -411
  88. package/docs/templates/agent-dag.supervised-implementation.json +620 -620
  89. package/docs/templates/backend-test-analysis.schema.json +44 -44
  90. package/docs/templates/backend-test-case-manifest.schema.json +190 -190
  91. package/docs/templates/backend-test-dag.classify.prompt.md +75 -75
  92. package/docs/templates/backend-test-dag.generate-pytest.prompt.md +204 -204
  93. package/docs/templates/backend-test-dag.json +559 -559
  94. package/docs/templates/backend-test-dag.retrospect.prompt.md +139 -139
  95. package/docs/templates/backend-test-dag.review-cases.prompt.md +83 -83
  96. package/docs/templates/backend-test-execution.schema.json +133 -133
  97. package/docs/templates/backend-test-result.schema.json +99 -99
  98. package/docs/templates/exec-plan.md +64 -64
  99. package/docs/templates/feature-spec.md +53 -53
  100. package/docs/templates/frontend-design-contract.md +42 -42
  101. package/docs/templates/frontend-eval/fixtures/failures/01-type-build-error.md +17 -17
  102. package/docs/templates/frontend-eval/fixtures/failures/02-unit-component-test-fail.md +16 -16
  103. package/docs/templates/frontend-eval/fixtures/failures/03-fixture-schema-drift.md +16 -16
  104. package/docs/templates/frontend-eval/fixtures/failures/04-missing-loading-empty-error-state.md +16 -16
  105. package/docs/templates/frontend-eval/fixtures/failures/05-forbidden-write-writeset-expansion.md +16 -16
  106. package/docs/templates/frontend-eval/fixtures/failures/06-unapproved-dependency-add.md +16 -16
  107. package/docs/templates/frontend-eval/fixtures/failures/07-mock-production-on.md +21 -21
  108. package/docs/templates/frontend-eval/fixtures/functional/01-simple-component-style.md +29 -29
  109. package/docs/templates/frontend-eval/fixtures/functional/02-form-validation.md +28 -28
  110. package/docs/templates/frontend-eval/fixtures/functional/03-list-detail-page.md +28 -28
  111. package/docs/templates/frontend-eval/fixtures/functional/04-api-mock.md +29 -29
  112. package/docs/templates/frontend-eval/fixtures/functional/05-permission-auth-gated-ui.md +27 -27
  113. package/docs/templates/frontend-eval/fixtures/functional/06-ssr-server-client-boundary.md +28 -28
  114. package/docs/templates/frontend-eval/fixtures/functional/07-shared-public-component-api.md +28 -28
  115. package/docs/templates/frontend-eval/fixtures/functional/08-pure-local-no-remote.md +27 -27
  116. package/docs/templates/frontend-eval/metrics.md +138 -138
  117. package/docs/templates/frontend-eval/smoke-targets.md +53 -53
  118. package/docs/templates/frontend-implementation-contract.schema.json +27 -27
  119. package/docs/templates/frontend-task-constraints.md +35 -35
  120. package/docs/templates/frontend-task-requirement.md +70 -70
  121. package/docs/templates/frontend-test-dag.generate-cases.prompt.md +5 -5
  122. package/docs/templates/frontend-test-dag.json +23 -23
  123. package/docs/templates/frontend-test-dag.retrieve-context.prompt.md +3 -3
  124. package/docs/templates/frontend-test-dag.retrospect.prompt.md +3 -3
  125. package/docs/templates/frontend-test-dag.review-cases.prompt.md +3 -3
  126. package/docs/templates/frontend-test-dag.review-execution.prompt.md +3 -3
  127. package/docs/templates/harness.schema.json +221 -221
  128. package/docs/templates/hybrid-dag.json +188 -188
  129. package/docs/templates/init-evolution-review.md +35 -35
  130. package/docs/templates/interactive-ui-round2-experiment.md +66 -66
  131. package/docs/templates/knowledge-graph-bootstrap-dag.json +118 -118
  132. package/docs/templates/knowledge-sync-dag.json +178 -178
  133. package/docs/templates/knowledge-sync-draft.schema.json +71 -71
  134. package/docs/templates/product-line/AGENTS.md +8 -8
  135. package/docs/templates/product-line/README.md +9 -9
  136. package/docs/templates/product-line/acceptance.yaml +14 -14
  137. package/docs/templates/product-line/closeout.yaml +9 -9
  138. package/docs/templates/product-line/design.md +13 -13
  139. package/docs/templates/product-line/links.md +10 -10
  140. package/docs/templates/product-line/requirement.md +17 -17
  141. package/docs/templates/product-line/task-graph.yaml +15 -15
  142. package/docs/templates/product-line/task.yaml +64 -64
  143. package/docs/templates/product-line/test-plan.md +7 -7
  144. package/docs/templates/production-readiness-checklist.md +57 -57
  145. package/docs/templates/progress-log.md +17 -17
  146. package/docs/templates/project-start-checklist.md +9 -9
  147. package/docs/templates/qa-report.md +48 -48
  148. package/docs/templates/sprint-contract.md +29 -29
  149. package/docs/templates/worker-dogfood-evidence.md +80 -80
  150. package/docs/templates/worker-dogfood-setup.md +68 -68
  151. package/docs/verification-matrix.md +70 -70
  152. package/examples/decision-gate-agent-dag.json +173 -173
  153. package/examples/example-dag.json +46 -46
  154. package/examples/hybrid-loop-agent-dag.json +188 -188
  155. package/harness.json +66 -66
  156. package/package.json +88 -52
  157. package/scripts/check-product-line-docs.sh +29 -29
  158. package/scripts/check-task-pool-root.sh +32 -32
  159. package/scripts/kb-bootstrap-init-skeleton.sh +240 -240
  160. package/scripts/kb-graph-incremental-prepare.mjs +386 -386
  161. package/scripts/kb-graph-incremental-prepare.sh +5 -5
  162. package/scripts/kb-graph-materialize.mjs +105 -105
  163. package/scripts/kb-graph-materialize.sh +4 -4
  164. package/scripts/kb-graph-promote.mjs +164 -164
  165. package/scripts/kb-graph-promote.sh +4 -4
  166. package/scripts/kb-query.mjs +554 -554
  167. package/scripts/kb-query.sh +5 -5
  168. package/skills/agent-worker/SKILL.md +39 -39
  169. package/skills/agent-worker/references/agent-worker-operator.md +60 -60
  170. package/skills/ai-engineering-context/SKILL.md +48 -48
  171. package/skills/analyze-product-dependencies/SKILL.md +67 -67
  172. package/skills/analyze-product-dependencies/agents/openai.yaml +4 -4
  173. package/skills/analyze-product-dependencies/references/api-documentation-schema.md +30 -30
  174. package/skills/analyze-product-dependencies/references/dependency-analysis-schema.md +28 -28
  175. package/skills/analyze-product-dependencies/references/example.md +76 -76
  176. package/skills/analyze-product-dependencies/references/forward-test-cases.md +35 -35
  177. package/skills/analyze-product-dependencies/references/input-contract.md +11 -11
  178. package/skills/analyze-product-dependencies/references/scouting-rules.md +61 -61
  179. package/skills/analyze-product-dependencies/scripts/test-validators.mjs +267 -267
  180. package/skills/analyze-product-dependencies/scripts/validate-api-documentation.mjs +101 -101
  181. package/skills/analyze-product-dependencies/scripts/validate-dependency-analysis.mjs +142 -142
  182. package/skills/analyze-product-dependencies/scripts/validate-product-requirement-input.mjs +76 -76
  183. package/skills/analyze-product-dependencies/scripts/validation-helpers.mjs +146 -146
  184. package/skills/analyze-product-requirements/SKILL.md +90 -90
  185. package/skills/analyze-product-requirements/agents/openai.yaml +4 -4
  186. package/skills/analyze-product-requirements/references/acceptance-criteria.md +91 -91
  187. package/skills/analyze-product-requirements/references/clarification-and-knowledge.md +56 -56
  188. package/skills/analyze-product-requirements/references/example.md +86 -86
  189. package/skills/analyze-product-requirements/references/forward-test-cases.md +66 -66
  190. package/skills/analyze-product-requirements/references/product-analysis-schema.md +32 -32
  191. package/skills/analyze-product-requirements/references/product-requirement-schema.md +33 -33
  192. package/skills/analyze-product-requirements/references/requirement-clarification-schema.md +35 -35
  193. package/skills/analyze-product-requirements/scripts/test-validators.mjs +193 -193
  194. package/skills/analyze-product-requirements/scripts/validate-product-analysis.mjs +69 -69
  195. package/skills/analyze-product-requirements/scripts/validate-product-requirement.mjs +97 -97
  196. package/skills/analyze-product-requirements/scripts/validate-requirement-clarification.mjs +98 -98
  197. package/skills/analyze-product-requirements/scripts/validation-helpers.mjs +156 -156
  198. package/skills/browser-tools/SKILL.md +196 -196
  199. package/skills/browser-tools/browser-content.js +103 -103
  200. package/skills/browser-tools/browser-cookies.js +35 -35
  201. package/skills/browser-tools/browser-eval.js +53 -53
  202. package/skills/browser-tools/browser-hn-scraper.js +108 -108
  203. package/skills/browser-tools/browser-nav.js +44 -44
  204. package/skills/browser-tools/browser-pick.js +162 -162
  205. package/skills/browser-tools/browser-screenshot.js +34 -34
  206. package/skills/browser-tools/browser-start.js +86 -86
  207. package/skills/browser-tools/package-lock.json +2556 -2556
  208. package/skills/browser-tools/package.json +19 -19
  209. package/skills/code-review-core/SKILL.md +20 -20
  210. package/skills/codebase-scout/SKILL.md +19 -19
  211. package/skills/frontend-design-review/SKILL.md +66 -66
  212. package/skills/frontend-design-review/references/review-checklist.md +58 -58
  213. package/skills/frontend-implementation/SKILL.md +49 -49
  214. package/skills/frontend-implementation/references/code-standards.md +32 -32
  215. package/skills/frontend-implementation/references/design-spec.md +46 -46
  216. package/skills/frontend-implementation/references/node-contracts.md +27 -27
  217. package/skills/frontend-review/SKILL.md +59 -59
  218. package/skills/frontend-review/references/review-findings.md +47 -47
  219. package/skills/frontend-verification/SKILL.md +53 -53
  220. package/skills/frontend-verification/references/verification-checklist.md +68 -68
  221. package/skills/grill-me/SKILL.md +10 -10
  222. package/skills/grill-with-docs/SKILL.md +88 -88
  223. package/skills/grill-with-docs/adr-format.md +47 -47
  224. package/skills/grill-with-docs/context-format.md +60 -60
  225. package/skills/init-capability-evolution/SKILL.md +70 -70
  226. package/skills/loop-agent/SKILL.md +151 -151
  227. package/skills/loop-agent/references/README.md +67 -67
  228. package/skills/loop-agent/references/command-reference.md +527 -527
  229. package/skills/loop-agent/references/docs-converge.md +126 -126
  230. package/skills/loop-agent/references/harness-policy.md +263 -263
  231. package/skills/loop-agent/references/hybrid-dag.md +243 -243
  232. package/skills/loop-agent/references/learned/README.md +21 -21
  233. package/skills/loop-agent/references/long-running-loop.md +57 -57
  234. package/skills/loop-agent/references/model-routing.md +36 -36
  235. package/skills/loop-agent/references/multi-worktree.md +54 -54
  236. package/skills/loop-agent/references/one-shot-runs.md +85 -85
  237. package/skills/loop-agent/references/orchestrator-and-interventions.md +169 -169
  238. package/skills/loop-agent/references/pi-prompt.md +23 -23
  239. package/skills/loop-agent/references/pi-subagent-assisted-mode.md +84 -84
  240. package/skills/loop-agent/references/post-implementation-and-patterns.md +44 -44
  241. package/skills/loop-agent/references/task-workflow.md +89 -89
  242. package/skills/loop-agent/references/verification-and-failure-handling.md +141 -141
  243. package/skills/playwright-cli/SKILL.md +420 -420
  244. package/skills/playwright-cli/references/element-attributes.md +23 -23
  245. package/skills/playwright-cli/references/playwright-tests.md +39 -39
  246. package/skills/playwright-cli/references/request-mocking.md +87 -87
  247. package/skills/playwright-cli/references/running-code.md +241 -241
  248. package/skills/playwright-cli/references/session-management.md +225 -225
  249. package/skills/playwright-cli/references/storage-state.md +275 -275
  250. package/skills/playwright-cli/references/test-generation.md +433 -433
  251. package/skills/playwright-cli/references/tracing.md +139 -139
  252. package/skills/playwright-cli/references/video-recording.md +143 -143
  253. package/skills/playwright-cli-case-generator/SKILL.md +74 -74
  254. package/skills/requesting-code-review/SKILL.md +101 -101
  255. package/skills/requesting-code-review/code-reviewer.md +168 -168
  256. package/skills/systematic-debugging/CREATION-LOG.md +119 -119
  257. package/skills/systematic-debugging/SKILL.md +296 -296
  258. package/skills/systematic-debugging/condition-based-waiting-example.ts +158 -158
  259. package/skills/systematic-debugging/condition-based-waiting.md +115 -115
  260. package/skills/systematic-debugging/defense-in-depth.md +122 -122
  261. package/skills/systematic-debugging/find-polluter.sh +63 -63
  262. package/skills/systematic-debugging/root-cause-tracing.md +169 -169
  263. package/skills/systematic-debugging/test-academic.md +14 -14
  264. package/skills/systematic-debugging/test-pressure-1.md +58 -58
  265. package/skills/systematic-debugging/test-pressure-2.md +68 -68
  266. package/skills/systematic-debugging/test-pressure-3.md +69 -69
  267. package/skills/test-driven-development/SKILL.md +20 -20
  268. package/skills/using-git-worktrees/SKILL.md +215 -215
  269. package/skills/verification-before-completion/SKILL.md +154 -154
  270. package/skills/webapp-testing/SKILL.md +19 -19
@@ -1,188 +1,188 @@
1
- {
2
- "version": 2,
3
- "title": "Hybrid DAG 示例:Pi 只读契约 + 并行侦察 + Pi 写入配置实现",
4
- "objective": "演示 v2 hybrid DAG:Pi 节点返回只读契约/审查结论,Pi writer 节点通过 toolProfile=write 执行受控独占实现,shell 节点负责确定性验证。",
5
- "successCriteria": [
6
- "contract-pi 返回可审计的只读实现契约摘要",
7
- "scout-src 与 scout-tests 并行只读侦察且互不写冲突",
8
- "implement-core 仅在声明的 writeSet 内修改代码",
9
- "review-pi 基于上游输出完成只读审查",
10
- "verify-shell 归档确定性验证命令输出",
11
- "dag validate 校验通过且 ranks 符合依赖拓扑"
12
- ],
13
- "globalConstraints": [
14
- "Do not commit runtime traces under .harness/dag-runs/",
15
- "Do not enable cross-node Pi runtime reuse by default",
16
- "Preserve CODE_AGENT_PI_BACKEND=cli-only rollback for Pi nodes",
17
- "Every task must explicitly declare executor; defaults.executor is schema-only and not a runtime fallback",
18
- "Prefer same-rank parallel read-only scouts over serial chains when outputs are independent",
19
- "Add depends_on only when a child truly needs upstream output; challenge single-chain topologies during review",
20
- "Same-rank exclusive writeSet entries must be disjoint",
21
- "exclusive implementer nodes must use narrow, concrete writeSet paths; never use ** or repo root",
22
- "Use Pi read-only scouts and Pi writer nodes with toolProfile=write; Cursor is only available through the explicit cursor-prompt sidecar",
23
- "Read-only nodes must not write root artifacts/**; root artifacts/ is reserved for explicit exclusive write nodes"
24
- ],
25
- "defaults": {
26
- "executor": "pi",
27
- "contextProfile": "slim",
28
- "skills": [
29
- "ai-engineering-context"
30
- ],
31
- "writePolicy": "read-only"
32
- },
33
- "skillsByRole": {
34
- "planner": [
35
- "loop-agent"
36
- ],
37
- "scout": [
38
- "ai-engineering-context"
39
- ],
40
- "implementer": [
41
- "verification-before-completion"
42
- ],
43
- "reviewer": [
44
- "requesting-code-review"
45
- ],
46
- "verifier": [
47
- "verification-before-completion",
48
- "systematic-debugging"
49
- ],
50
- "closeout": [
51
- "loop-agent",
52
- "verification-before-completion"
53
- ]
54
- },
55
- "executorModels": {
56
- "pi": {
57
- "LOW": "gpt-5.3-codex-spark",
58
- "MED": "glm-5.2",
59
- "HIGH": "gpt-5.5"
60
- }
61
- },
62
- "tasks": [
63
- {
64
- "id": "contract-pi",
65
- "depends_on": [],
66
- "role": "planner",
67
- "executor": "pi",
68
- "complexity": "MED",
69
- "writePolicy": "read-only",
70
- "allowedPaths": [
71
- "./**",
72
- "docs/**"
73
- ],
74
- "forbiddenPaths": [
75
- ".harness/**",
76
- "artifacts/**"
77
- ],
78
- "outputContract": "Plain Markdown implementation contract summary; no file writes.",
79
- "subtask_prompt": "阅读 ./README.md 与 docs/agent-dag-runner.md,返回 10 行以内的实现契约摘要(只读分析 + 文档建议,不改代码/文档/artifacts)。指出是否适合并行 scout 与窄 writeSet。"
80
- },
81
- {
82
- "id": "scout-src",
83
- "depends_on": [
84
- "contract-pi"
85
- ],
86
- "role": "scout",
87
- "executor": "pi",
88
- "complexity": "LOW",
89
- "writePolicy": "read-only",
90
- "allowedPaths": [
91
- "./src/workflows/dag/**"
92
- ],
93
- "forbiddenPaths": [
94
- ".harness/**",
95
- "artifacts/**"
96
- ],
97
- "outputContract": "Plain Markdown reconnaissance summary of DAG-related source modules; no file writes.",
98
- "subtask_prompt": "只读审计 ./src/workflows/dag/ 中与 DAG 相关的模块,列出关键文件与职责(不要改文件)。"
99
- },
100
- {
101
- "id": "scout-tests",
102
- "depends_on": [
103
- "contract-pi"
104
- ],
105
- "role": "scout",
106
- "executor": "pi",
107
- "complexity": "LOW",
108
- "writePolicy": "read-only",
109
- "allowedPaths": [
110
- "./test/**"
111
- ],
112
- "forbiddenPaths": [
113
- ".harness/**",
114
- "artifacts/**"
115
- ],
116
- "outputContract": "Plain Markdown test coverage gap summary; no file writes.",
117
- "subtask_prompt": "只读审计 ./test/ 中与 dag-runner 相关的测试覆盖,列出缺口建议(不要改文件)。"
118
- },
119
- {
120
- "id": "implement-core",
121
- "depends_on": [
122
- "scout-src",
123
- "scout-tests"
124
- ],
125
- "role": "implementer",
126
- "executor": "pi",
127
- "complexity": "HIGH",
128
- "writePolicy": "exclusive",
129
- "writeSet": [
130
- "./src/workflows/dag/runner.ts"
131
- ],
132
- "allowedPaths": [
133
- "./src/workflows/dag/**"
134
- ],
135
- "forbiddenPaths": [
136
- ".harness/**",
137
- "artifacts/**"
138
- ],
139
- "outputContract": "Implementation summary with changed files, tests run, and residual risks.",
140
- "subtask_prompt": "基于上游侦察结果,在 writeSet 范围内做最小必要改动并说明风险(若无需改动则输出 no-op 理由)。",
141
- "toolProfile": "write"
142
- },
143
- {
144
- "id": "review-pi",
145
- "depends_on": [
146
- "implement-core"
147
- ],
148
- "role": "reviewer",
149
- "executor": "pi",
150
- "complexity": "MED",
151
- "writePolicy": "read-only",
152
- "allowedPaths": [
153
- "./**",
154
- "docs/**"
155
- ],
156
- "forbiddenPaths": [
157
- ".harness/**",
158
- "artifacts/**"
159
- ],
160
- "outputContract": "Plain Markdown review summary; no file writes.",
161
- "subtask_prompt": "审查上游实现与侦察结论是否一致,列出范围漂移、验证缺口与残余风险(只读,不改文件/文档/artifacts)。"
162
- },
163
- {
164
- "id": "verify-shell",
165
- "depends_on": [
166
- "implement-core"
167
- ],
168
- "role": "verifier",
169
- "executor": "shell",
170
- "complexity": "LOW",
171
- "writePolicy": "read-only",
172
- "allowedPaths": [
173
- "./**"
174
- ],
175
- "forbiddenPaths": [
176
- ".harness/**",
177
- "artifacts/**"
178
- ],
179
- "outputContract": "Archived shell verification stdout/stderr with exit codes; no worktree writes.",
180
- "subtask_prompt": "Run deterministic loop-agent verification and archive command outputs.",
181
- "shell": {
182
- "preset": "loop-agent-standard-verify",
183
- "cwd": ".",
184
- "timeoutMs": 300000
185
- }
186
- }
187
- ]
188
- }
1
+ {
2
+ "version": 2,
3
+ "title": "Hybrid DAG 示例:Pi 只读契约 + 并行侦察 + Pi 写入配置实现",
4
+ "objective": "演示 v2 hybrid DAG:Pi 节点返回只读契约/审查结论,Pi writer 节点通过 toolProfile=write 执行受控独占实现,shell 节点负责确定性验证。",
5
+ "successCriteria": [
6
+ "contract-pi 返回可审计的只读实现契约摘要",
7
+ "scout-src 与 scout-tests 并行只读侦察且互不写冲突",
8
+ "implement-core 仅在声明的 writeSet 内修改代码",
9
+ "review-pi 基于上游输出完成只读审查",
10
+ "verify-shell 归档确定性验证命令输出",
11
+ "dag validate 校验通过且 ranks 符合依赖拓扑"
12
+ ],
13
+ "globalConstraints": [
14
+ "Do not commit runtime traces under .harness/dag-runs/",
15
+ "Do not enable cross-node Pi runtime reuse by default",
16
+ "Preserve CODE_AGENT_PI_BACKEND=cli-only rollback for Pi nodes",
17
+ "Every task must explicitly declare executor; defaults.executor is schema-only and not a runtime fallback",
18
+ "Prefer same-rank parallel read-only scouts over serial chains when outputs are independent",
19
+ "Add depends_on only when a child truly needs upstream output; challenge single-chain topologies during review",
20
+ "Same-rank exclusive writeSet entries must be disjoint",
21
+ "exclusive implementer nodes must use narrow, concrete writeSet paths; never use ** or repo root",
22
+ "Use Pi read-only scouts and Pi writer nodes with toolProfile=write; Cursor is only available through the explicit cursor-prompt sidecar",
23
+ "Read-only nodes must not write root artifacts/**; root artifacts/ is reserved for explicit exclusive write nodes"
24
+ ],
25
+ "defaults": {
26
+ "executor": "pi",
27
+ "contextProfile": "slim",
28
+ "skills": [
29
+ "ai-engineering-context"
30
+ ],
31
+ "writePolicy": "read-only"
32
+ },
33
+ "skillsByRole": {
34
+ "planner": [
35
+ "loop-agent"
36
+ ],
37
+ "scout": [
38
+ "ai-engineering-context"
39
+ ],
40
+ "implementer": [
41
+ "verification-before-completion"
42
+ ],
43
+ "reviewer": [
44
+ "requesting-code-review"
45
+ ],
46
+ "verifier": [
47
+ "verification-before-completion",
48
+ "systematic-debugging"
49
+ ],
50
+ "closeout": [
51
+ "loop-agent",
52
+ "verification-before-completion"
53
+ ]
54
+ },
55
+ "executorModels": {
56
+ "pi": {
57
+ "LOW": "gpt-5.3-codex-spark",
58
+ "MED": "glm-5.2",
59
+ "HIGH": "gpt-5.5"
60
+ }
61
+ },
62
+ "tasks": [
63
+ {
64
+ "id": "contract-pi",
65
+ "depends_on": [],
66
+ "role": "planner",
67
+ "executor": "pi",
68
+ "complexity": "MED",
69
+ "writePolicy": "read-only",
70
+ "allowedPaths": [
71
+ "./**",
72
+ "docs/**"
73
+ ],
74
+ "forbiddenPaths": [
75
+ ".harness/**",
76
+ "artifacts/**"
77
+ ],
78
+ "outputContract": "Plain Markdown implementation contract summary; no file writes.",
79
+ "subtask_prompt": "阅读 ./README.md 与 docs/agent-dag-runner.md,返回 10 行以内的实现契约摘要(只读分析 + 文档建议,不改代码/文档/artifacts)。指出是否适合并行 scout 与窄 writeSet。"
80
+ },
81
+ {
82
+ "id": "scout-src",
83
+ "depends_on": [
84
+ "contract-pi"
85
+ ],
86
+ "role": "scout",
87
+ "executor": "pi",
88
+ "complexity": "LOW",
89
+ "writePolicy": "read-only",
90
+ "allowedPaths": [
91
+ "./src/workflows/dag/**"
92
+ ],
93
+ "forbiddenPaths": [
94
+ ".harness/**",
95
+ "artifacts/**"
96
+ ],
97
+ "outputContract": "Plain Markdown reconnaissance summary of DAG-related source modules; no file writes.",
98
+ "subtask_prompt": "只读审计 ./src/workflows/dag/ 中与 DAG 相关的模块,列出关键文件与职责(不要改文件)。"
99
+ },
100
+ {
101
+ "id": "scout-tests",
102
+ "depends_on": [
103
+ "contract-pi"
104
+ ],
105
+ "role": "scout",
106
+ "executor": "pi",
107
+ "complexity": "LOW",
108
+ "writePolicy": "read-only",
109
+ "allowedPaths": [
110
+ "./test/**"
111
+ ],
112
+ "forbiddenPaths": [
113
+ ".harness/**",
114
+ "artifacts/**"
115
+ ],
116
+ "outputContract": "Plain Markdown test coverage gap summary; no file writes.",
117
+ "subtask_prompt": "只读审计 ./test/ 中与 dag-runner 相关的测试覆盖,列出缺口建议(不要改文件)。"
118
+ },
119
+ {
120
+ "id": "implement-core",
121
+ "depends_on": [
122
+ "scout-src",
123
+ "scout-tests"
124
+ ],
125
+ "role": "implementer",
126
+ "executor": "pi",
127
+ "complexity": "HIGH",
128
+ "writePolicy": "exclusive",
129
+ "writeSet": [
130
+ "./src/workflows/dag/runner.ts"
131
+ ],
132
+ "allowedPaths": [
133
+ "./src/workflows/dag/**"
134
+ ],
135
+ "forbiddenPaths": [
136
+ ".harness/**",
137
+ "artifacts/**"
138
+ ],
139
+ "outputContract": "Implementation summary with changed files, tests run, and residual risks.",
140
+ "subtask_prompt": "基于上游侦察结果,在 writeSet 范围内做最小必要改动并说明风险(若无需改动则输出 no-op 理由)。",
141
+ "toolProfile": "write"
142
+ },
143
+ {
144
+ "id": "review-pi",
145
+ "depends_on": [
146
+ "implement-core"
147
+ ],
148
+ "role": "reviewer",
149
+ "executor": "pi",
150
+ "complexity": "MED",
151
+ "writePolicy": "read-only",
152
+ "allowedPaths": [
153
+ "./**",
154
+ "docs/**"
155
+ ],
156
+ "forbiddenPaths": [
157
+ ".harness/**",
158
+ "artifacts/**"
159
+ ],
160
+ "outputContract": "Plain Markdown review summary; no file writes.",
161
+ "subtask_prompt": "审查上游实现与侦察结论是否一致,列出范围漂移、验证缺口与残余风险(只读,不改文件/文档/artifacts)。"
162
+ },
163
+ {
164
+ "id": "verify-shell",
165
+ "depends_on": [
166
+ "implement-core"
167
+ ],
168
+ "role": "verifier",
169
+ "executor": "shell",
170
+ "complexity": "LOW",
171
+ "writePolicy": "read-only",
172
+ "allowedPaths": [
173
+ "./**"
174
+ ],
175
+ "forbiddenPaths": [
176
+ ".harness/**",
177
+ "artifacts/**"
178
+ ],
179
+ "outputContract": "Archived shell verification stdout/stderr with exit codes; no worktree writes.",
180
+ "subtask_prompt": "Run deterministic loop-agent verification and archive command outputs.",
181
+ "shell": {
182
+ "preset": "loop-agent-standard-verify",
183
+ "cwd": ".",
184
+ "timeoutMs": 300000
185
+ }
186
+ }
187
+ ]
188
+ }
@@ -1,35 +1,35 @@
1
- # Init Evolution Review
2
-
3
- Date:
4
- Base: `<full commit SHA or resolvable --base ref>`
5
- Head: `<full commit SHA for the reviewed HEAD>`
6
-
7
- `bash scripts/check-init-evolution-needed.sh --strict --base <ref>` 接受 Base 精确匹配该 `<ref>`、Head 为当前 `HEAD` 或其可解析祖先提交的报告。历史报告、`working tree` 等不可解析文字不能为其他变更范围放行严格检查。当 Head 是祖先时,`reportHead..HEAD` 区间内出现新的 `model-review` 高影响路径会使报告失效;仅 `advisory` 或 `surface-check` 的后续变化不影响已完成的高影响审查。
8
-
9
- ## Changed Surface
10
-
11
- -
12
-
13
- ## Decision
14
-
15
- Choose one:
16
-
17
- - No init impact
18
- - Surface check only
19
- - Init update required
20
-
21
- Rationale:
22
-
23
- ## Updates Made
24
-
25
- -
26
-
27
- ## Verification
28
-
29
- ```bash
30
- # commands and results
31
- ```
32
-
33
- ## Residual Risk
34
-
35
- -
1
+ # Init Evolution Review
2
+
3
+ Date:
4
+ Base: `<full commit SHA or resolvable --base ref>`
5
+ Head: `<full commit SHA for the reviewed HEAD>`
6
+
7
+ `bash scripts/check-init-evolution-needed.sh --strict --base <ref>` 接受 Base 精确匹配该 `<ref>`、Head 为当前 `HEAD` 或其可解析祖先提交的报告。历史报告、`working tree` 等不可解析文字不能为其他变更范围放行严格检查。当 Head 是祖先时,`reportHead..HEAD` 区间内出现新的 `model-review` 高影响路径会使报告失效;仅 `advisory` 或 `surface-check` 的后续变化不影响已完成的高影响审查。
8
+
9
+ ## Changed Surface
10
+
11
+ -
12
+
13
+ ## Decision
14
+
15
+ Choose one:
16
+
17
+ - No init impact
18
+ - Surface check only
19
+ - Init update required
20
+
21
+ Rationale:
22
+
23
+ ## Updates Made
24
+
25
+ -
26
+
27
+ ## Verification
28
+
29
+ ```bash
30
+ # commands and results
31
+ ```
32
+
33
+ ## Residual Risk
34
+
35
+ -
@@ -1,66 +1,66 @@
1
- # Interactive UI Round-2 A/B/C Experiment
2
-
3
- ## Frozen inputs
4
-
5
- - Target repo / commit:
6
- - Controller version:
7
- - TaskSpec / acceptance hash:
8
- - Base DAG:
9
- - Model matrix:
10
-
11
- Prepare the fixture and three DAGs after installing a published controller that contains `interactive-ui` support:
12
-
13
- ```bash
14
- bash scripts-local/setup-drill-round2-react.sh /tmp/drill-round2-react-target
15
- npm run build
16
- node scripts-local/prepare-round2-ui-experiment.mjs \
17
- /tmp/drill-round2-react-target \
18
- features/F-2026-001/tasks/FE-001.yaml \
19
- /tmp/round2-ui-experiment \
20
- loop-agent
21
- node scripts-local/build-round2-ui-variants.mjs \
22
- /tmp/round2-ui-experiment/round2-ui-base.json \
23
- /tmp/round2-ui-experiment/variants
24
- ```
25
-
26
- Validate every generated variant with the same frozen controller before running it. Use a unique run ID for A, B, and C; do not rewrite `executorModels`.
27
-
28
- ## Variants
29
-
30
- | Variant | Writer prompt | Writer tier | Run ID | Result |
31
- | --- | --- | --- | --- | --- |
32
- | A | interactive-ui contract | MED | | |
33
- | B | default contract | HIGH | | |
34
- | C | interactive-ui contract | HIGH | | |
35
-
36
- ## Metrics
37
-
38
- | Metric | A | B | C |
39
- | --- | --- | --- | --- |
40
- | First-pass review gate pass | | | |
41
- | Framework-native `.tsx` component | | | |
42
- | Route/parent integration | | | |
43
- | DOM interaction tests | | | |
44
- | Helper-only escape | | | |
45
- | Duration | | | |
46
- | Tokens | | | |
47
-
48
- ## Required evidence per run
49
-
50
- - DAG JSON and run ID
51
- - implement/repair writer model and prompt profile
52
- - changed component path
53
- - integration path
54
- - interaction test path
55
- - DOM assertions mapped to AC-FE-001
56
- - review verdict and deterministic test output
57
- - diff boundary audit
58
-
59
- ## Decision
60
-
61
- - Production default:
62
- - Evidence:
63
- - Cost/quality trade-off:
64
- - Follow-up:
65
-
66
- Do not conclude from a single run when provider or environment failures occurred. Re-run the affected variant with the same frozen inputs and a new run ID.
1
+ # Interactive UI Round-2 A/B/C Experiment
2
+
3
+ ## Frozen inputs
4
+
5
+ - Target repo / commit:
6
+ - Controller version:
7
+ - TaskSpec / acceptance hash:
8
+ - Base DAG:
9
+ - Model matrix:
10
+
11
+ Prepare the fixture and three DAGs after installing a published controller that contains `interactive-ui` support:
12
+
13
+ ```bash
14
+ bash scripts-local/setup-drill-round2-react.sh /tmp/drill-round2-react-target
15
+ npm run build
16
+ node scripts-local/prepare-round2-ui-experiment.mjs \
17
+ /tmp/drill-round2-react-target \
18
+ features/F-2026-001/tasks/FE-001.yaml \
19
+ /tmp/round2-ui-experiment \
20
+ loop-agent
21
+ node scripts-local/build-round2-ui-variants.mjs \
22
+ /tmp/round2-ui-experiment/round2-ui-base.json \
23
+ /tmp/round2-ui-experiment/variants
24
+ ```
25
+
26
+ Validate every generated variant with the same frozen controller before running it. Use a unique run ID for A, B, and C; do not rewrite `executorModels`.
27
+
28
+ ## Variants
29
+
30
+ | Variant | Writer prompt | Writer tier | Run ID | Result |
31
+ | --- | --- | --- | --- | --- |
32
+ | A | interactive-ui contract | MED | | |
33
+ | B | default contract | HIGH | | |
34
+ | C | interactive-ui contract | HIGH | | |
35
+
36
+ ## Metrics
37
+
38
+ | Metric | A | B | C |
39
+ | --- | --- | --- | --- |
40
+ | First-pass review gate pass | | | |
41
+ | Framework-native `.tsx` component | | | |
42
+ | Route/parent integration | | | |
43
+ | DOM interaction tests | | | |
44
+ | Helper-only escape | | | |
45
+ | Duration | | | |
46
+ | Tokens | | | |
47
+
48
+ ## Required evidence per run
49
+
50
+ - DAG JSON and run ID
51
+ - implement/repair writer model and prompt profile
52
+ - changed component path
53
+ - integration path
54
+ - interaction test path
55
+ - DOM assertions mapped to AC-FE-001
56
+ - review verdict and deterministic test output
57
+ - diff boundary audit
58
+
59
+ ## Decision
60
+
61
+ - Production default:
62
+ - Evidence:
63
+ - Cost/quality trade-off:
64
+ - Follow-up:
65
+
66
+ Do not conclude from a single run when provider or environment failures occurred. Re-run the affected variant with the same frozen inputs and a new run ID.