@tea-agent/loop-agent 0.5.0 → 0.6.0

This diff represents the content of publicly available package versions that have been released to one of the supported registries. The information contained in this diff is provided for informational purposes only and reflects changes between package versions as they appear in their respective public registries.
Files changed (133) hide show
  1. package/AGENTS.md +142 -142
  2. package/CHANGELOG.md +116 -98
  3. package/README.md +195 -195
  4. package/bin/agent-worker.js +22 -22
  5. package/bin/loop-agent.js +21 -21
  6. package/dist/application/dag/args.js +9 -1
  7. package/dist/application/dag/run-dag.js +16 -2
  8. package/dist/cli/command-definitions.js +22 -4
  9. package/dist/cli/help.js +3 -2
  10. package/dist/cli/program.js +7 -5
  11. package/dist/commands/import-prd.js +76 -0
  12. package/dist/commands/init.js +467 -457
  13. package/dist/commands/instructions.js +90 -58
  14. package/dist/commands/loop-benchmark.js +11 -11
  15. package/dist/commands/pi-reuse-benchmark.js +16 -16
  16. package/dist/executors/cursor-executor.js +1 -1
  17. package/dist/executors/dag-pi-executor.js +1 -0
  18. package/dist/executors/pi-sdk-executor.js +63 -1
  19. package/dist/shared/preview.js +39 -0
  20. package/dist/task/runtime.js +27 -27
  21. package/dist/task/source-references.js +221 -0
  22. package/dist/worker/cli.js +62 -1
  23. package/dist/worker/loop-agent/loop-agent-client.js +97 -5
  24. package/dist/worker/materialize/harness-task-materializer.js +162 -5
  25. package/dist/worker/observability/event-store.js +82 -0
  26. package/dist/worker/observability/events.js +79 -0
  27. package/dist/worker/observability/progress-composite.js +33 -0
  28. package/dist/worker/observability/read-model.js +1013 -0
  29. package/dist/worker/observability/snapshot-store.js +43 -0
  30. package/dist/worker/observability/types.js +1 -0
  31. package/dist/worker/observe/paths.js +64 -0
  32. package/dist/worker/observe/routes.js +423 -0
  33. package/dist/worker/observe/server.js +61 -0
  34. package/dist/worker/observe/static/app.js +1419 -0
  35. package/dist/worker/observe/static/index.html +63 -0
  36. package/dist/worker/observe/static/styles.css +613 -0
  37. package/dist/worker/pool/failure-routing.js +41 -6
  38. package/dist/worker/pool/run-store.js +50 -0
  39. package/dist/worker/progress-reporter.js +0 -18
  40. package/dist/worker/run-task/run-task.js +327 -92
  41. package/dist/worker/runner/run-ready.js +112 -4
  42. package/dist/workflows/dag/canvas-observer.js +275 -275
  43. package/dist/workflows/dag/event-observer.js +132 -0
  44. package/dist/workflows/dag/init-hybrid.js +146 -13
  45. package/dist/workflows/dag/observer-compose.js +52 -0
  46. package/docs/README.md +74 -72
  47. package/docs/agent-dag-recovery-playbook.md +184 -184
  48. package/docs/agent-dag-runner.md +42 -42
  49. package/docs/architecture/runtime-boundaries.md +162 -147
  50. package/docs/cursor-executor-usage.md +25 -25
  51. package/docs/decisions/README.md +3 -3
  52. package/docs/design/README.md +49 -36
  53. package/docs/development-principles.md +73 -73
  54. package/docs/dynamic-workflow-dag-engine-roadmap.md +1749 -1749
  55. package/docs/exec-plans/README.md +6 -6
  56. package/docs/exec-plans/active/README.md +12 -7
  57. package/docs/exec-plans/completed/README.md +31 -19
  58. package/docs/feature-workflow.md +186 -186
  59. package/docs/harness-methodology-debugging.md +153 -153
  60. package/docs/harness-methodology-tdd.md +130 -130
  61. package/docs/harness-methodology-verification.md +27 -27
  62. package/docs/init-surface.manifest.json +205 -199
  63. package/docs/loop-agent-harness.md +55 -42
  64. package/docs/production-readiness.md +96 -96
  65. package/docs/progress/README.md +3 -3
  66. package/docs/reports/README.md +9 -5
  67. package/docs/skills/README.md +6 -6
  68. package/docs/skills/vetted-skill-registry.md +26 -26
  69. package/docs/templates/adr.md +60 -60
  70. package/docs/templates/agent-dag-authority-surface-audit.prompt.md +94 -94
  71. package/docs/templates/agent-dag-decision-envelope.schema.json +213 -213
  72. package/docs/templates/agent-dag-decision-gate-dogfood-report.md +117 -117
  73. package/docs/templates/agent-dag-decision-gate.prompt.md +246 -246
  74. package/docs/templates/agent-dag-process-supervisor.prompt.md +98 -98
  75. package/docs/templates/agent-dag-report.schema.json +454 -454
  76. package/docs/templates/agent-dag-review-verdict.prompt.md +68 -68
  77. package/docs/templates/agent-dag.base.json +195 -195
  78. package/docs/templates/agent-dag.final-verification.json +190 -190
  79. package/docs/templates/agent-dag.schema.json +316 -316
  80. package/docs/templates/agent-dag.supervised-implementation.json +500 -500
  81. package/docs/templates/exec-plan.md +64 -64
  82. package/docs/templates/feature-spec.md +53 -53
  83. package/docs/templates/hybrid-dag.json +193 -193
  84. package/docs/templates/init-evolution-review.md +33 -33
  85. package/docs/templates/production-readiness-checklist.md +57 -57
  86. package/docs/templates/progress-log.md +17 -17
  87. package/docs/templates/project-start-checklist.md +9 -9
  88. package/docs/templates/qa-report.md +48 -48
  89. package/docs/templates/sprint-contract.md +29 -29
  90. package/docs/templates/worker-dogfood-evidence.md +52 -0
  91. package/docs/templates/worker-dogfood-setup.md +48 -0
  92. package/docs/verification-matrix.md +41 -41
  93. package/examples/decision-gate-agent-dag.json +123 -123
  94. package/examples/example-dag.json +51 -51
  95. package/examples/hybrid-loop-agent-dag.json +194 -194
  96. package/harness.json +70 -69
  97. package/package.json +66 -66
  98. package/skills/ai-engineering-context/SKILL.md +48 -48
  99. package/skills/code-review-core/SKILL.md +20 -20
  100. package/skills/codebase-scout/SKILL.md +19 -19
  101. package/skills/init-capability-evolution/SKILL.md +69 -69
  102. package/skills/loop-agent/SKILL.md +149 -147
  103. package/skills/loop-agent/references/README.md +67 -67
  104. package/skills/loop-agent/references/command-reference.md +412 -403
  105. package/skills/loop-agent/references/harness-policy.md +263 -259
  106. package/skills/loop-agent/references/hybrid-dag.md +216 -216
  107. package/skills/loop-agent/references/learned/README.md +21 -21
  108. package/skills/loop-agent/references/long-running-loop.md +59 -59
  109. package/skills/loop-agent/references/model-routing.md +36 -36
  110. package/skills/loop-agent/references/multi-worktree.md +54 -54
  111. package/skills/loop-agent/references/one-shot-runs.md +85 -85
  112. package/skills/loop-agent/references/orchestrator-and-interventions.md +169 -169
  113. package/skills/loop-agent/references/pi-prompt.md +23 -23
  114. package/skills/loop-agent/references/pi-subagent-assisted-mode.md +81 -81
  115. package/skills/loop-agent/references/post-implementation-and-patterns.md +44 -44
  116. package/skills/loop-agent/references/task-workflow.md +89 -84
  117. package/skills/loop-agent/references/verification-and-failure-handling.md +128 -128
  118. package/skills/requesting-code-review/SKILL.md +101 -101
  119. package/skills/requesting-code-review/code-reviewer.md +168 -168
  120. package/skills/systematic-debugging/CREATION-LOG.md +119 -119
  121. package/skills/systematic-debugging/SKILL.md +296 -296
  122. package/skills/systematic-debugging/condition-based-waiting-example.ts +158 -158
  123. package/skills/systematic-debugging/condition-based-waiting.md +115 -115
  124. package/skills/systematic-debugging/defense-in-depth.md +122 -122
  125. package/skills/systematic-debugging/find-polluter.sh +63 -63
  126. package/skills/systematic-debugging/root-cause-tracing.md +169 -169
  127. package/skills/systematic-debugging/test-academic.md +14 -14
  128. package/skills/systematic-debugging/test-pressure-1.md +58 -58
  129. package/skills/systematic-debugging/test-pressure-2.md +68 -68
  130. package/skills/systematic-debugging/test-pressure-3.md +69 -69
  131. package/skills/test-driven-development/SKILL.md +20 -20
  132. package/skills/verification-before-completion/SKILL.md +154 -154
  133. package/skills/webapp-testing/SKILL.md +19 -19
@@ -1,119 +1,119 @@
1
- # Creation Log: Systematic Debugging Skill
2
-
3
- 提取、结构化与 bulletproofing 关键 skill 的 reference example。
4
-
5
- ## Source Material
6
-
7
- 从 `~/.claude/CLAUDE.md` 提取 debugging framework:
8
- - 4-phase systematic process(Investigation → Pattern Analysis → Hypothesis → Implementation)
9
- - Core mandate:ALWAYS find root cause,NEVER fix symptoms
10
- - 设计以 resist time pressure 与 rationalization 的规则
11
-
12
- ## Extraction Decisions
13
-
14
- **What to include:**
15
- - 完整 4-phase framework 及所有 rules
16
- - Anti-shortcuts("NEVER fix symptom"、"STOP and re-analyze")
17
- - Pressure-resistant language("even if faster"、"even if I seem in a hurry")
18
- - 各 phase 的 concrete steps
19
-
20
- **What to leave out:**
21
- - Project-specific context
22
- - 同一 rule 的 repetitive variations
23
- - Narrative explanations(condensed 为 principles)
24
-
25
- ## Structure Following skill-creation/SKILL.md
26
-
27
- 1. **Rich when_to_use** — 含 symptoms 与 anti-patterns
28
- 2. **Type: technique** — 带 steps 的 concrete process
29
- 3. **Keywords** — "root cause"、"symptom"、"workaround"、"debugging"、"investigation"
30
- 4. **Flowchart** — "fix failed" 决策点 → re-analyze vs add more fixes
31
- 5. **Phase-by-phase breakdown** — Scannable checklist format
32
- 6. **Anti-patterns section** — 什么 NOT to do(对本 skill 关键)
33
-
34
- ## Bulletproofing Elements
35
-
36
- Framework 设计以 resist rationalization under pressure:
37
-
38
- ### Language Choices
39
- - "ALWAYS" / "NEVER"(非 "should" / "try to")
40
- - "even if faster" / "even if I seem in a hurry"
41
- - "STOP and re-analyze"(explicit pause)
42
- - "Don't skip past"(捕获 actual behavior)
43
-
44
- ### Structural Defenses
45
- - **Phase 1 required** — 不能 skip to implementation
46
- - **Single hypothesis rule** — 强制思考,防止 shotgun fixes
47
- - **Explicit failure mode** — "IF your first fix doesn't work" 及 mandatory action
48
- - **Anti-patterns section** — 展示 shortcuts 的确切样子
49
-
50
- ### Redundancy
51
- - Root cause mandate 在 overview + when_to_use + Phase 1 + implementation rules
52
- - "NEVER fix symptom" 在不同 contexts 出现 4 次
53
- - 各 phase 有 explicit "don't skip" guidance
54
-
55
- ## Testing Approach
56
-
57
- 按 skills/meta/testing-skills-with-subagents 创建 4 个 validation tests:
58
-
59
- ### Test 1: Academic Context (No Pressure)
60
- - Simple bug,无 time pressure
61
- - **Result:** Perfect compliance,complete investigation
62
-
63
- ### Test 2: Time Pressure + Obvious Quick Fix
64
- - User "in a hurry",symptom fix 看起来 easy
65
- - **Result:** Resisted shortcut,followed full process,found real root cause
66
-
67
- ### Test 3: Complex System + Uncertainty
68
- - Multi-layer failure, unclear 能否 find root cause
69
- - **Result:** Systematic investigation,traced through all layers,found source
70
-
71
- ### Test 4: Failed First Fix
72
- - Hypothesis 无效,temptation 加 more fixes
73
- - **Result:** Stopped,re-analyzed,formed new hypothesis(no shotgun)
74
-
75
- **All tests passed.** No rationalizations found.
76
-
77
- ## Iterations
78
-
79
- ### Initial Version
80
- - Complete 4-phase framework
81
- - Anti-patterns section
82
- - Flowchart for "fix failed" decision
83
-
84
- ### Enhancement 1: TDD Reference
85
- - Added link to skills/testing/test-driven-development
86
- - Note explaining TDD's "simplest code" ≠ debugging's "root cause"
87
- - Prevents confusion between methodologies
88
-
89
- ## Final Outcome
90
-
91
- Bulletproof skill that:
92
- - ✅ Clearly mandates root cause investigation
93
- - ✅ Resists time pressure rationalization
94
- - ✅ Provides concrete steps for each phase
95
- - ✅ Shows anti-patterns explicitly
96
- - ✅ Tested under multiple pressure scenarios
97
- - ✅ Clarifies relationship to TDD
98
- - ✅ Ready for use
99
-
100
- ## Key Insight
101
-
102
- **Most important bulletproofing:** Anti-patterns section 展示 moment 里 feel justified 的 exact shortcuts。当 Claude 想 "I'll just add this one quick fix",看到 listed as wrong 的 exact pattern 产生 cognitive friction。
103
-
104
- ## Usage Example
105
-
106
- 遇到 bug 时:
107
- 1. Load skill: skills/debugging/systematic-debugging
108
- 2. Read overview (10 sec) — reminded of mandate
109
- 3. Follow Phase 1 checklist — forced investigation
110
- 4. If tempted to skip — see anti-pattern,stop
111
- 5. Complete all phases — root cause found
112
-
113
- **Time investment:** 5-10 minutes
114
- **Time saved:** Hours of symptom-whack-a-mole
115
-
116
- ---
117
-
118
- *Created: 2025-10-03*
119
- *Purpose: Reference example for skill extraction and bulletproofing*
1
+ # Creation Log: Systematic Debugging Skill
2
+
3
+ 提取、结构化与 bulletproofing 关键 skill 的 reference example。
4
+
5
+ ## Source Material
6
+
7
+ 从 `~/.claude/CLAUDE.md` 提取 debugging framework:
8
+ - 4-phase systematic process(Investigation → Pattern Analysis → Hypothesis → Implementation)
9
+ - Core mandate:ALWAYS find root cause,NEVER fix symptoms
10
+ - 设计以 resist time pressure 与 rationalization 的规则
11
+
12
+ ## Extraction Decisions
13
+
14
+ **What to include:**
15
+ - 完整 4-phase framework 及所有 rules
16
+ - Anti-shortcuts("NEVER fix symptom"、"STOP and re-analyze")
17
+ - Pressure-resistant language("even if faster"、"even if I seem in a hurry")
18
+ - 各 phase 的 concrete steps
19
+
20
+ **What to leave out:**
21
+ - Project-specific context
22
+ - 同一 rule 的 repetitive variations
23
+ - Narrative explanations(condensed 为 principles)
24
+
25
+ ## Structure Following skill-creation/SKILL.md
26
+
27
+ 1. **Rich when_to_use** — 含 symptoms 与 anti-patterns
28
+ 2. **Type: technique** — 带 steps 的 concrete process
29
+ 3. **Keywords** — "root cause"、"symptom"、"workaround"、"debugging"、"investigation"
30
+ 4. **Flowchart** — "fix failed" 决策点 → re-analyze vs add more fixes
31
+ 5. **Phase-by-phase breakdown** — Scannable checklist format
32
+ 6. **Anti-patterns section** — 什么 NOT to do(对本 skill 关键)
33
+
34
+ ## Bulletproofing Elements
35
+
36
+ Framework 设计以 resist rationalization under pressure:
37
+
38
+ ### Language Choices
39
+ - "ALWAYS" / "NEVER"(非 "should" / "try to")
40
+ - "even if faster" / "even if I seem in a hurry"
41
+ - "STOP and re-analyze"(explicit pause)
42
+ - "Don't skip past"(捕获 actual behavior)
43
+
44
+ ### Structural Defenses
45
+ - **Phase 1 required** — 不能 skip to implementation
46
+ - **Single hypothesis rule** — 强制思考,防止 shotgun fixes
47
+ - **Explicit failure mode** — "IF your first fix doesn't work" 及 mandatory action
48
+ - **Anti-patterns section** — 展示 shortcuts 的确切样子
49
+
50
+ ### Redundancy
51
+ - Root cause mandate 在 overview + when_to_use + Phase 1 + implementation rules
52
+ - "NEVER fix symptom" 在不同 contexts 出现 4 次
53
+ - 各 phase 有 explicit "don't skip" guidance
54
+
55
+ ## Testing Approach
56
+
57
+ 按 skills/meta/testing-skills-with-subagents 创建 4 个 validation tests:
58
+
59
+ ### Test 1: Academic Context (No Pressure)
60
+ - Simple bug,无 time pressure
61
+ - **Result:** Perfect compliance,complete investigation
62
+
63
+ ### Test 2: Time Pressure + Obvious Quick Fix
64
+ - User "in a hurry",symptom fix 看起来 easy
65
+ - **Result:** Resisted shortcut,followed full process,found real root cause
66
+
67
+ ### Test 3: Complex System + Uncertainty
68
+ - Multi-layer failure, unclear 能否 find root cause
69
+ - **Result:** Systematic investigation,traced through all layers,found source
70
+
71
+ ### Test 4: Failed First Fix
72
+ - Hypothesis 无效,temptation 加 more fixes
73
+ - **Result:** Stopped,re-analyzed,formed new hypothesis(no shotgun)
74
+
75
+ **All tests passed.** No rationalizations found.
76
+
77
+ ## Iterations
78
+
79
+ ### Initial Version
80
+ - Complete 4-phase framework
81
+ - Anti-patterns section
82
+ - Flowchart for "fix failed" decision
83
+
84
+ ### Enhancement 1: TDD Reference
85
+ - Added link to skills/testing/test-driven-development
86
+ - Note explaining TDD's "simplest code" ≠ debugging's "root cause"
87
+ - Prevents confusion between methodologies
88
+
89
+ ## Final Outcome
90
+
91
+ Bulletproof skill that:
92
+ - ✅ Clearly mandates root cause investigation
93
+ - ✅ Resists time pressure rationalization
94
+ - ✅ Provides concrete steps for each phase
95
+ - ✅ Shows anti-patterns explicitly
96
+ - ✅ Tested under multiple pressure scenarios
97
+ - ✅ Clarifies relationship to TDD
98
+ - ✅ Ready for use
99
+
100
+ ## Key Insight
101
+
102
+ **Most important bulletproofing:** Anti-patterns section 展示 moment 里 feel justified 的 exact shortcuts。当 Claude 想 "I'll just add this one quick fix",看到 listed as wrong 的 exact pattern 产生 cognitive friction。
103
+
104
+ ## Usage Example
105
+
106
+ 遇到 bug 时:
107
+ 1. Load skill: skills/debugging/systematic-debugging
108
+ 2. Read overview (10 sec) — reminded of mandate
109
+ 3. Follow Phase 1 checklist — forced investigation
110
+ 4. If tempted to skip — see anti-pattern,stop
111
+ 5. Complete all phases — root cause found
112
+
113
+ **Time investment:** 5-10 minutes
114
+ **Time saved:** Hours of symptom-whack-a-mole
115
+
116
+ ---
117
+
118
+ *Created: 2025-10-03*
119
+ *Purpose: Reference example for skill extraction and bulletproofing*