@tea-agent/loop-agent 0.7.5 → 0.8.0

This diff represents the content of publicly available package versions that have been released to one of the supported registries. The information contained in this diff is provided for informational purposes only and reflects changes between package versions as they appear in their respective public registries.
Files changed (127) hide show
  1. package/AGENTS.md +143 -142
  2. package/CHANGELOG.md +148 -164
  3. package/README.md +206 -204
  4. package/bin/agent-worker.js +22 -22
  5. package/bin/loop-agent.js +21 -21
  6. package/dist/commands/init.js +518 -488
  7. package/dist/commands/loop-benchmark.js +11 -11
  8. package/dist/commands/pi-reuse-benchmark.js +16 -16
  9. package/dist/executors/cursor-executor.js +1 -1
  10. package/dist/governance/manifest-types.js +1 -1
  11. package/dist/task/runtime.js +27 -27
  12. package/dist/worker/cli.js +3 -3
  13. package/dist/worker/observability/event-store.js +2 -1
  14. package/dist/worker/observability/read-model.js +13 -11
  15. package/dist/worker/observe/paths.js +2 -2
  16. package/dist/worker/observe/routes.js +4 -3
  17. package/dist/worker/observe/static/app.js +1479 -1480
  18. package/dist/worker/observe/static/dag-layout.d.ts +31 -31
  19. package/dist/worker/observe/static/dag-layout.js +83 -83
  20. package/dist/worker/observe/static/index.html +63 -63
  21. package/dist/worker/observe/static/styles.css +722 -722
  22. package/dist/worker/pool/run-store.js +7 -8
  23. package/dist/worker/run-task/run-task.js +11 -2
  24. package/dist/worker/runner/run-ready.js +1 -1
  25. package/dist/workflows/dag/canvas-observer.js +275 -275
  26. package/docs/README.md +80 -79
  27. package/docs/agent-dag-recovery-playbook.md +184 -184
  28. package/docs/agent-dag-runner.md +42 -42
  29. package/docs/architecture/runtime-boundaries.md +162 -162
  30. package/docs/cursor-executor-usage.md +25 -25
  31. package/docs/decisions/README.md +3 -3
  32. package/docs/design/README.md +49 -49
  33. package/docs/development-principles.md +73 -73
  34. package/docs/dynamic-workflow-dag-engine-roadmap.md +1749 -1749
  35. package/docs/exec-plans/README.md +6 -6
  36. package/docs/exec-plans/active/README.md +11 -11
  37. package/docs/exec-plans/completed/README.md +35 -34
  38. package/docs/feature-workflow.md +187 -187
  39. package/docs/harness-methodology-debugging.md +153 -153
  40. package/docs/harness-methodology-tdd.md +130 -130
  41. package/docs/harness-methodology-verification.md +27 -27
  42. package/docs/init-surface.manifest.json +245 -241
  43. package/docs/loop-agent-harness.md +63 -55
  44. package/docs/production-readiness.md +96 -96
  45. package/docs/progress/README.md +3 -3
  46. package/docs/reports/README.md +9 -9
  47. package/docs/skills/README.md +6 -6
  48. package/docs/skills/vetted-skill-registry.md +26 -26
  49. package/docs/templates/adr.md +60 -60
  50. package/docs/templates/agent-dag-authority-surface-audit.prompt.md +94 -94
  51. package/docs/templates/agent-dag-decision-envelope.schema.json +213 -213
  52. package/docs/templates/agent-dag-decision-gate-dogfood-report.md +117 -117
  53. package/docs/templates/agent-dag-decision-gate.prompt.md +246 -246
  54. package/docs/templates/agent-dag-process-supervisor.prompt.md +98 -98
  55. package/docs/templates/agent-dag-report.schema.json +454 -454
  56. package/docs/templates/agent-dag-review-verdict.prompt.md +68 -68
  57. package/docs/templates/agent-dag.base.json +195 -195
  58. package/docs/templates/agent-dag.final-verification.json +190 -190
  59. package/docs/templates/agent-dag.schema.json +316 -316
  60. package/docs/templates/agent-dag.supervised-implementation.json +500 -500
  61. package/docs/templates/exec-plan.md +64 -64
  62. package/docs/templates/feature-spec.md +53 -53
  63. package/docs/templates/harness.schema.json +218 -0
  64. package/docs/templates/hybrid-dag.json +193 -193
  65. package/docs/templates/init-evolution-review.md +33 -33
  66. package/docs/templates/interactive-ui-round2-experiment.md +66 -66
  67. package/docs/templates/product-line/AGENTS.md +8 -8
  68. package/docs/templates/product-line/README.md +9 -9
  69. package/docs/templates/product-line/acceptance.yaml +14 -14
  70. package/docs/templates/product-line/closeout.yaml +9 -9
  71. package/docs/templates/product-line/design.md +13 -13
  72. package/docs/templates/product-line/links.md +10 -10
  73. package/docs/templates/product-line/requirement.md +17 -17
  74. package/docs/templates/product-line/task-graph.yaml +15 -15
  75. package/docs/templates/product-line/task.yaml +65 -65
  76. package/docs/templates/product-line/test-plan.md +7 -7
  77. package/docs/templates/production-readiness-checklist.md +57 -57
  78. package/docs/templates/progress-log.md +17 -17
  79. package/docs/templates/project-start-checklist.md +9 -9
  80. package/docs/templates/qa-report.md +48 -48
  81. package/docs/templates/sprint-contract.md +29 -29
  82. package/docs/templates/worker-dogfood-evidence.md +52 -52
  83. package/docs/templates/worker-dogfood-setup.md +48 -48
  84. package/docs/verification-matrix.md +49 -49
  85. package/examples/decision-gate-agent-dag.json +123 -123
  86. package/examples/example-dag.json +51 -51
  87. package/examples/hybrid-loop-agent-dag.json +194 -194
  88. package/harness.json +73 -71
  89. package/package.json +68 -67
  90. package/scripts/check-product-line-docs.sh +22 -22
  91. package/scripts/check-task-pool-root.sh +32 -0
  92. package/skills/ai-engineering-context/SKILL.md +48 -48
  93. package/skills/code-review-core/SKILL.md +20 -20
  94. package/skills/codebase-scout/SKILL.md +19 -19
  95. package/skills/init-capability-evolution/SKILL.md +69 -69
  96. package/skills/loop-agent/SKILL.md +149 -149
  97. package/skills/loop-agent/references/README.md +67 -67
  98. package/skills/loop-agent/references/command-reference.md +432 -412
  99. package/skills/loop-agent/references/harness-policy.md +263 -263
  100. package/skills/loop-agent/references/hybrid-dag.md +216 -216
  101. package/skills/loop-agent/references/learned/README.md +21 -21
  102. package/skills/loop-agent/references/long-running-loop.md +59 -59
  103. package/skills/loop-agent/references/model-routing.md +36 -36
  104. package/skills/loop-agent/references/multi-worktree.md +54 -54
  105. package/skills/loop-agent/references/one-shot-runs.md +85 -85
  106. package/skills/loop-agent/references/orchestrator-and-interventions.md +169 -169
  107. package/skills/loop-agent/references/pi-prompt.md +23 -23
  108. package/skills/loop-agent/references/pi-subagent-assisted-mode.md +81 -81
  109. package/skills/loop-agent/references/post-implementation-and-patterns.md +44 -44
  110. package/skills/loop-agent/references/task-workflow.md +89 -89
  111. package/skills/loop-agent/references/verification-and-failure-handling.md +128 -128
  112. package/skills/requesting-code-review/SKILL.md +101 -101
  113. package/skills/requesting-code-review/code-reviewer.md +168 -168
  114. package/skills/systematic-debugging/CREATION-LOG.md +119 -119
  115. package/skills/systematic-debugging/SKILL.md +296 -296
  116. package/skills/systematic-debugging/condition-based-waiting-example.ts +158 -158
  117. package/skills/systematic-debugging/condition-based-waiting.md +115 -115
  118. package/skills/systematic-debugging/defense-in-depth.md +122 -122
  119. package/skills/systematic-debugging/find-polluter.sh +63 -63
  120. package/skills/systematic-debugging/root-cause-tracing.md +169 -169
  121. package/skills/systematic-debugging/test-academic.md +14 -14
  122. package/skills/systematic-debugging/test-pressure-1.md +58 -58
  123. package/skills/systematic-debugging/test-pressure-2.md +68 -68
  124. package/skills/systematic-debugging/test-pressure-3.md +69 -69
  125. package/skills/test-driven-development/SKILL.md +20 -20
  126. package/skills/verification-before-completion/SKILL.md +154 -154
  127. package/skills/webapp-testing/SKILL.md +19 -19
@@ -1,119 +1,119 @@
1
- # Creation Log: Systematic Debugging Skill
2
-
3
- 提取、结构化与 bulletproofing 关键 skill 的 reference example。
4
-
5
- ## Source Material
6
-
7
- 从 `~/.claude/CLAUDE.md` 提取 debugging framework:
8
- - 4-phase systematic process(Investigation → Pattern Analysis → Hypothesis → Implementation)
9
- - Core mandate:ALWAYS find root cause,NEVER fix symptoms
10
- - 设计以 resist time pressure 与 rationalization 的规则
11
-
12
- ## Extraction Decisions
13
-
14
- **What to include:**
15
- - 完整 4-phase framework 及所有 rules
16
- - Anti-shortcuts("NEVER fix symptom"、"STOP and re-analyze")
17
- - Pressure-resistant language("even if faster"、"even if I seem in a hurry")
18
- - 各 phase 的 concrete steps
19
-
20
- **What to leave out:**
21
- - Project-specific context
22
- - 同一 rule 的 repetitive variations
23
- - Narrative explanations(condensed 为 principles)
24
-
25
- ## Structure Following skill-creation/SKILL.md
26
-
27
- 1. **Rich when_to_use** — 含 symptoms 与 anti-patterns
28
- 2. **Type: technique** — 带 steps 的 concrete process
29
- 3. **Keywords** — "root cause"、"symptom"、"workaround"、"debugging"、"investigation"
30
- 4. **Flowchart** — "fix failed" 决策点 → re-analyze vs add more fixes
31
- 5. **Phase-by-phase breakdown** — Scannable checklist format
32
- 6. **Anti-patterns section** — 什么 NOT to do(对本 skill 关键)
33
-
34
- ## Bulletproofing Elements
35
-
36
- Framework 设计以 resist rationalization under pressure:
37
-
38
- ### Language Choices
39
- - "ALWAYS" / "NEVER"(非 "should" / "try to")
40
- - "even if faster" / "even if I seem in a hurry"
41
- - "STOP and re-analyze"(explicit pause)
42
- - "Don't skip past"(捕获 actual behavior)
43
-
44
- ### Structural Defenses
45
- - **Phase 1 required** — 不能 skip to implementation
46
- - **Single hypothesis rule** — 强制思考,防止 shotgun fixes
47
- - **Explicit failure mode** — "IF your first fix doesn't work" 及 mandatory action
48
- - **Anti-patterns section** — 展示 shortcuts 的确切样子
49
-
50
- ### Redundancy
51
- - Root cause mandate 在 overview + when_to_use + Phase 1 + implementation rules
52
- - "NEVER fix symptom" 在不同 contexts 出现 4 次
53
- - 各 phase 有 explicit "don't skip" guidance
54
-
55
- ## Testing Approach
56
-
57
- 按 skills/meta/testing-skills-with-subagents 创建 4 个 validation tests:
58
-
59
- ### Test 1: Academic Context (No Pressure)
60
- - Simple bug,无 time pressure
61
- - **Result:** Perfect compliance,complete investigation
62
-
63
- ### Test 2: Time Pressure + Obvious Quick Fix
64
- - User "in a hurry",symptom fix 看起来 easy
65
- - **Result:** Resisted shortcut,followed full process,found real root cause
66
-
67
- ### Test 3: Complex System + Uncertainty
68
- - Multi-layer failure, unclear 能否 find root cause
69
- - **Result:** Systematic investigation,traced through all layers,found source
70
-
71
- ### Test 4: Failed First Fix
72
- - Hypothesis 无效,temptation 加 more fixes
73
- - **Result:** Stopped,re-analyzed,formed new hypothesis(no shotgun)
74
-
75
- **All tests passed.** No rationalizations found.
76
-
77
- ## Iterations
78
-
79
- ### Initial Version
80
- - Complete 4-phase framework
81
- - Anti-patterns section
82
- - Flowchart for "fix failed" decision
83
-
84
- ### Enhancement 1: TDD Reference
85
- - Added link to skills/testing/test-driven-development
86
- - Note explaining TDD's "simplest code" ≠ debugging's "root cause"
87
- - Prevents confusion between methodologies
88
-
89
- ## Final Outcome
90
-
91
- Bulletproof skill that:
92
- - ✅ Clearly mandates root cause investigation
93
- - ✅ Resists time pressure rationalization
94
- - ✅ Provides concrete steps for each phase
95
- - ✅ Shows anti-patterns explicitly
96
- - ✅ Tested under multiple pressure scenarios
97
- - ✅ Clarifies relationship to TDD
98
- - ✅ Ready for use
99
-
100
- ## Key Insight
101
-
102
- **Most important bulletproofing:** Anti-patterns section 展示 moment 里 feel justified 的 exact shortcuts。当 Claude 想 "I'll just add this one quick fix",看到 listed as wrong 的 exact pattern 产生 cognitive friction。
103
-
104
- ## Usage Example
105
-
106
- 遇到 bug 时:
107
- 1. Load skill: skills/debugging/systematic-debugging
108
- 2. Read overview (10 sec) — reminded of mandate
109
- 3. Follow Phase 1 checklist — forced investigation
110
- 4. If tempted to skip — see anti-pattern,stop
111
- 5. Complete all phases — root cause found
112
-
113
- **Time investment:** 5-10 minutes
114
- **Time saved:** Hours of symptom-whack-a-mole
115
-
116
- ---
117
-
118
- *Created: 2025-10-03*
119
- *Purpose: Reference example for skill extraction and bulletproofing*
1
+ # Creation Log: Systematic Debugging Skill
2
+
3
+ 提取、结构化与 bulletproofing 关键 skill 的 reference example。
4
+
5
+ ## Source Material
6
+
7
+ 从 `~/.claude/CLAUDE.md` 提取 debugging framework:
8
+ - 4-phase systematic process(Investigation → Pattern Analysis → Hypothesis → Implementation)
9
+ - Core mandate:ALWAYS find root cause,NEVER fix symptoms
10
+ - 设计以 resist time pressure 与 rationalization 的规则
11
+
12
+ ## Extraction Decisions
13
+
14
+ **What to include:**
15
+ - 完整 4-phase framework 及所有 rules
16
+ - Anti-shortcuts("NEVER fix symptom"、"STOP and re-analyze")
17
+ - Pressure-resistant language("even if faster"、"even if I seem in a hurry")
18
+ - 各 phase 的 concrete steps
19
+
20
+ **What to leave out:**
21
+ - Project-specific context
22
+ - 同一 rule 的 repetitive variations
23
+ - Narrative explanations(condensed 为 principles)
24
+
25
+ ## Structure Following skill-creation/SKILL.md
26
+
27
+ 1. **Rich when_to_use** — 含 symptoms 与 anti-patterns
28
+ 2. **Type: technique** — 带 steps 的 concrete process
29
+ 3. **Keywords** — "root cause"、"symptom"、"workaround"、"debugging"、"investigation"
30
+ 4. **Flowchart** — "fix failed" 决策点 → re-analyze vs add more fixes
31
+ 5. **Phase-by-phase breakdown** — Scannable checklist format
32
+ 6. **Anti-patterns section** — 什么 NOT to do(对本 skill 关键)
33
+
34
+ ## Bulletproofing Elements
35
+
36
+ Framework 设计以 resist rationalization under pressure:
37
+
38
+ ### Language Choices
39
+ - "ALWAYS" / "NEVER"(非 "should" / "try to")
40
+ - "even if faster" / "even if I seem in a hurry"
41
+ - "STOP and re-analyze"(explicit pause)
42
+ - "Don't skip past"(捕获 actual behavior)
43
+
44
+ ### Structural Defenses
45
+ - **Phase 1 required** — 不能 skip to implementation
46
+ - **Single hypothesis rule** — 强制思考,防止 shotgun fixes
47
+ - **Explicit failure mode** — "IF your first fix doesn't work" 及 mandatory action
48
+ - **Anti-patterns section** — 展示 shortcuts 的确切样子
49
+
50
+ ### Redundancy
51
+ - Root cause mandate 在 overview + when_to_use + Phase 1 + implementation rules
52
+ - "NEVER fix symptom" 在不同 contexts 出现 4 次
53
+ - 各 phase 有 explicit "don't skip" guidance
54
+
55
+ ## Testing Approach
56
+
57
+ 按 skills/meta/testing-skills-with-subagents 创建 4 个 validation tests:
58
+
59
+ ### Test 1: Academic Context (No Pressure)
60
+ - Simple bug,无 time pressure
61
+ - **Result:** Perfect compliance,complete investigation
62
+
63
+ ### Test 2: Time Pressure + Obvious Quick Fix
64
+ - User "in a hurry",symptom fix 看起来 easy
65
+ - **Result:** Resisted shortcut,followed full process,found real root cause
66
+
67
+ ### Test 3: Complex System + Uncertainty
68
+ - Multi-layer failure, unclear 能否 find root cause
69
+ - **Result:** Systematic investigation,traced through all layers,found source
70
+
71
+ ### Test 4: Failed First Fix
72
+ - Hypothesis 无效,temptation 加 more fixes
73
+ - **Result:** Stopped,re-analyzed,formed new hypothesis(no shotgun)
74
+
75
+ **All tests passed.** No rationalizations found.
76
+
77
+ ## Iterations
78
+
79
+ ### Initial Version
80
+ - Complete 4-phase framework
81
+ - Anti-patterns section
82
+ - Flowchart for "fix failed" decision
83
+
84
+ ### Enhancement 1: TDD Reference
85
+ - Added link to skills/testing/test-driven-development
86
+ - Note explaining TDD's "simplest code" ≠ debugging's "root cause"
87
+ - Prevents confusion between methodologies
88
+
89
+ ## Final Outcome
90
+
91
+ Bulletproof skill that:
92
+ - ✅ Clearly mandates root cause investigation
93
+ - ✅ Resists time pressure rationalization
94
+ - ✅ Provides concrete steps for each phase
95
+ - ✅ Shows anti-patterns explicitly
96
+ - ✅ Tested under multiple pressure scenarios
97
+ - ✅ Clarifies relationship to TDD
98
+ - ✅ Ready for use
99
+
100
+ ## Key Insight
101
+
102
+ **Most important bulletproofing:** Anti-patterns section 展示 moment 里 feel justified 的 exact shortcuts。当 Claude 想 "I'll just add this one quick fix",看到 listed as wrong 的 exact pattern 产生 cognitive friction。
103
+
104
+ ## Usage Example
105
+
106
+ 遇到 bug 时:
107
+ 1. Load skill: skills/debugging/systematic-debugging
108
+ 2. Read overview (10 sec) — reminded of mandate
109
+ 3. Follow Phase 1 checklist — forced investigation
110
+ 4. If tempted to skip — see anti-pattern,stop
111
+ 5. Complete all phases — root cause found
112
+
113
+ **Time investment:** 5-10 minutes
114
+ **Time saved:** Hours of symptom-whack-a-mole
115
+
116
+ ---
117
+
118
+ *Created: 2025-10-03*
119
+ *Purpose: Reference example for skill extraction and bulletproofing*