@tea-agent/loop-agent 0.5.0 → 0.7.0

This diff represents the content of publicly available package versions that have been released to one of the supported registries. The information contained in this diff is provided for informational purposes only and reflects changes between package versions as they appear in their respective public registries.
Files changed (136) hide show
  1. package/AGENTS.md +142 -142
  2. package/CHANGELOG.md +132 -98
  3. package/README.md +195 -195
  4. package/bin/agent-worker.js +22 -22
  5. package/bin/loop-agent.js +21 -21
  6. package/dist/application/dag/args.js +9 -1
  7. package/dist/application/dag/run-dag.js +16 -2
  8. package/dist/cli/command-definitions.js +22 -4
  9. package/dist/cli/help.js +3 -2
  10. package/dist/cli/program.js +7 -5
  11. package/dist/commands/import-prd.js +76 -0
  12. package/dist/commands/init.js +467 -457
  13. package/dist/commands/instructions.js +90 -58
  14. package/dist/commands/loop-benchmark.js +11 -11
  15. package/dist/commands/pi-reuse-benchmark.js +16 -16
  16. package/dist/executors/cursor-executor.js +1 -1
  17. package/dist/executors/dag-pi-executor.js +1 -0
  18. package/dist/executors/pi-sdk-executor.js +63 -1
  19. package/dist/shared/preview.js +39 -0
  20. package/dist/task/config-types.js +3 -0
  21. package/dist/task/runtime.js +27 -27
  22. package/dist/task/source-references.js +221 -0
  23. package/dist/worker/cli.js +62 -1
  24. package/dist/worker/loop-agent/loop-agent-client.js +97 -5
  25. package/dist/worker/materialize/harness-task-materializer.js +166 -5
  26. package/dist/worker/observability/event-store.js +82 -0
  27. package/dist/worker/observability/events.js +79 -0
  28. package/dist/worker/observability/progress-composite.js +33 -0
  29. package/dist/worker/observability/read-model.js +1013 -0
  30. package/dist/worker/observability/snapshot-store.js +43 -0
  31. package/dist/worker/observability/types.js +1 -0
  32. package/dist/worker/observe/paths.js +64 -0
  33. package/dist/worker/observe/routes.js +423 -0
  34. package/dist/worker/observe/server.js +61 -0
  35. package/dist/worker/observe/static/app.js +1419 -0
  36. package/dist/worker/observe/static/index.html +63 -0
  37. package/dist/worker/observe/static/styles.css +613 -0
  38. package/dist/worker/pool/failure-routing.js +41 -6
  39. package/dist/worker/pool/run-store.js +50 -0
  40. package/dist/worker/progress-reporter.js +0 -18
  41. package/dist/worker/run-task/run-task.js +327 -92
  42. package/dist/worker/runner/run-ready.js +112 -4
  43. package/dist/worker/task-spec/schema.js +2 -1
  44. package/dist/workflows/dag/canvas-observer.js +275 -275
  45. package/dist/workflows/dag/event-observer.js +132 -0
  46. package/dist/workflows/dag/init-hybrid.js +182 -21
  47. package/dist/workflows/dag/observer-compose.js +52 -0
  48. package/docs/README.md +75 -72
  49. package/docs/agent-dag-recovery-playbook.md +184 -184
  50. package/docs/agent-dag-runner.md +42 -42
  51. package/docs/architecture/runtime-boundaries.md +162 -147
  52. package/docs/cursor-executor-usage.md +25 -25
  53. package/docs/decisions/README.md +3 -3
  54. package/docs/design/README.md +49 -36
  55. package/docs/development-principles.md +73 -73
  56. package/docs/dynamic-workflow-dag-engine-roadmap.md +1749 -1749
  57. package/docs/exec-plans/README.md +6 -6
  58. package/docs/exec-plans/active/README.md +12 -7
  59. package/docs/exec-plans/completed/README.md +32 -19
  60. package/docs/feature-workflow.md +186 -186
  61. package/docs/harness-methodology-debugging.md +153 -153
  62. package/docs/harness-methodology-tdd.md +130 -130
  63. package/docs/harness-methodology-verification.md +27 -27
  64. package/docs/init-surface.manifest.json +208 -199
  65. package/docs/loop-agent-harness.md +55 -42
  66. package/docs/production-readiness.md +96 -96
  67. package/docs/progress/README.md +3 -3
  68. package/docs/reports/README.md +9 -5
  69. package/docs/skills/README.md +6 -6
  70. package/docs/skills/vetted-skill-registry.md +26 -26
  71. package/docs/templates/adr.md +60 -60
  72. package/docs/templates/agent-dag-authority-surface-audit.prompt.md +94 -94
  73. package/docs/templates/agent-dag-decision-envelope.schema.json +213 -213
  74. package/docs/templates/agent-dag-decision-gate-dogfood-report.md +117 -117
  75. package/docs/templates/agent-dag-decision-gate.prompt.md +246 -246
  76. package/docs/templates/agent-dag-process-supervisor.prompt.md +98 -98
  77. package/docs/templates/agent-dag-report.schema.json +454 -454
  78. package/docs/templates/agent-dag-review-verdict.prompt.md +68 -68
  79. package/docs/templates/agent-dag.base.json +195 -195
  80. package/docs/templates/agent-dag.final-verification.json +190 -190
  81. package/docs/templates/agent-dag.schema.json +316 -316
  82. package/docs/templates/agent-dag.supervised-implementation.json +500 -500
  83. package/docs/templates/exec-plan.md +64 -64
  84. package/docs/templates/feature-spec.md +53 -53
  85. package/docs/templates/hybrid-dag.json +193 -193
  86. package/docs/templates/init-evolution-review.md +33 -33
  87. package/docs/templates/interactive-ui-round2-experiment.md +66 -0
  88. package/docs/templates/production-readiness-checklist.md +57 -57
  89. package/docs/templates/progress-log.md +17 -17
  90. package/docs/templates/project-start-checklist.md +9 -9
  91. package/docs/templates/qa-report.md +48 -48
  92. package/docs/templates/sprint-contract.md +29 -29
  93. package/docs/templates/worker-dogfood-evidence.md +52 -0
  94. package/docs/templates/worker-dogfood-setup.md +48 -0
  95. package/docs/verification-matrix.md +41 -41
  96. package/examples/decision-gate-agent-dag.json +123 -123
  97. package/examples/example-dag.json +51 -51
  98. package/examples/hybrid-loop-agent-dag.json +194 -194
  99. package/harness.json +70 -69
  100. package/package.json +66 -66
  101. package/skills/ai-engineering-context/SKILL.md +48 -48
  102. package/skills/code-review-core/SKILL.md +20 -20
  103. package/skills/codebase-scout/SKILL.md +19 -19
  104. package/skills/init-capability-evolution/SKILL.md +69 -69
  105. package/skills/loop-agent/SKILL.md +149 -147
  106. package/skills/loop-agent/references/README.md +67 -67
  107. package/skills/loop-agent/references/command-reference.md +412 -403
  108. package/skills/loop-agent/references/harness-policy.md +263 -259
  109. package/skills/loop-agent/references/hybrid-dag.md +216 -216
  110. package/skills/loop-agent/references/learned/README.md +21 -21
  111. package/skills/loop-agent/references/long-running-loop.md +59 -59
  112. package/skills/loop-agent/references/model-routing.md +36 -36
  113. package/skills/loop-agent/references/multi-worktree.md +54 -54
  114. package/skills/loop-agent/references/one-shot-runs.md +85 -85
  115. package/skills/loop-agent/references/orchestrator-and-interventions.md +169 -169
  116. package/skills/loop-agent/references/pi-prompt.md +23 -23
  117. package/skills/loop-agent/references/pi-subagent-assisted-mode.md +81 -81
  118. package/skills/loop-agent/references/post-implementation-and-patterns.md +44 -44
  119. package/skills/loop-agent/references/task-workflow.md +89 -84
  120. package/skills/loop-agent/references/verification-and-failure-handling.md +128 -128
  121. package/skills/requesting-code-review/SKILL.md +101 -101
  122. package/skills/requesting-code-review/code-reviewer.md +168 -168
  123. package/skills/systematic-debugging/CREATION-LOG.md +119 -119
  124. package/skills/systematic-debugging/SKILL.md +296 -296
  125. package/skills/systematic-debugging/condition-based-waiting-example.ts +158 -158
  126. package/skills/systematic-debugging/condition-based-waiting.md +115 -115
  127. package/skills/systematic-debugging/defense-in-depth.md +122 -122
  128. package/skills/systematic-debugging/find-polluter.sh +63 -63
  129. package/skills/systematic-debugging/root-cause-tracing.md +169 -169
  130. package/skills/systematic-debugging/test-academic.md +14 -14
  131. package/skills/systematic-debugging/test-pressure-1.md +58 -58
  132. package/skills/systematic-debugging/test-pressure-2.md +68 -68
  133. package/skills/systematic-debugging/test-pressure-3.md +69 -69
  134. package/skills/test-driven-development/SKILL.md +20 -20
  135. package/skills/verification-before-completion/SKILL.md +154 -154
  136. package/skills/webapp-testing/SKILL.md +19 -19
@@ -1,101 +1,101 @@
1
- ---
2
- name: requesting-code-review
3
- description: 在完成任务、实现 major features,或 merge 前验证 work 是否满足 requirements 时使用
4
- ---
5
-
6
- # Requesting Code Review
7
-
8
- Dispatch code reviewer subagent,在问题级联前捕获 issue。Reviewer 获得精确 crafted 的 evaluation context — 绝不是你的 session history。这使 reviewer 聚焦 work product,而非你的 thought process,并保留你自己的 context 以继续工作。
9
-
10
- **Core principle:** Review early, review often.
11
-
12
- ## When to Request Review
13
-
14
- **Mandatory:**
15
- - subagent-driven development 中每个 task 之后
16
- - 完成 major feature 之后
17
- - merge 到 main 之前
18
-
19
- **Optional but valuable:**
20
- - 卡住时(fresh perspective)
21
- - refactoring 前(baseline check)
22
- - 修复 complex bug 之后
23
-
24
- ## How to Request
25
-
26
- **1. Get git SHAs:**
27
- ```bash
28
- BASE_SHA=$(git rev-parse HEAD~1) # or origin/main
29
- HEAD_SHA=$(git rev-parse HEAD)
30
- ```
31
-
32
- **2. Use the code reviewer template**(本 skill 目录下的 `code-reviewer.md`):
33
-
34
- **Placeholders:**
35
- - `{DESCRIPTION}` — 简要 summary of what you built
36
- - `{PLAN_OR_REQUIREMENTS}` — 它应做什么(contract、exec plan 或 requirements)
37
- - `{BASE_SHA}` — Starting commit
38
- - `{HEAD_SHA}` — Ending commit
39
-
40
- **3. Act on feedback:**
41
- - Critical issues 立即修复
42
- - Important issues 在继续前修复
43
- - Minor issues 稍后处理
44
- - Reviewer 有误时 push back(附 reasoning)
45
-
46
- ## Example
47
-
48
- ```
49
- [Just completed Task 2: Add verification function]
50
-
51
- You: Let me request code review before proceeding.
52
-
53
- BASE_SHA=$(git log --oneline | grep "Task 1" | head -1 | awk '{print $1}')
54
- HEAD_SHA=$(git rev-parse HEAD)
55
-
56
- [Dispatch code reviewer subagent]
57
- DESCRIPTION: Added verifyIndex() and repairIndex() with 4 issue types
58
- PLAN_OR_REQUIREMENTS: Task 2 from docs/exec-plans/active/deployment-plan.md
59
- BASE_SHA: a7981ec
60
- HEAD_SHA: 3df7661
61
-
62
- [Subagent returns]:
63
- Strengths: Clean architecture, real tests
64
- Issues:
65
- Important: Missing progress indicators
66
- Minor: Magic number (100) for reporting interval
67
- Assessment: Ready to proceed
68
-
69
- You: [Fix progress indicators]
70
- [Continue to Task 3]
71
- ```
72
-
73
- ## Integration with Harness Workflow
74
-
75
- **每个 work chunk 之后(Plan → Contract → Implement → Verify → Handoff):**
76
- - Implement 之后、Verify 之前 review
77
- - 在问题 compound 前捕获
78
- - 进入 next task 前修复
79
-
80
- **Before merge / Handoff:**
81
- - 宣称 complete 前 review
82
- - 对照 contract acceptance criteria 验证
83
-
84
- **Ad-Hoc Development:**
85
- - merge 前 review
86
- - 卡住时 review
87
-
88
- ## Red Flags
89
-
90
- **Never:**
91
- - 因 "it's simple" 跳过 review
92
- - 忽略 Critical issues
93
- - 带着未修复的 Important issues 继续
94
- - 与 valid technical feedback 争辩
95
-
96
- **If reviewer wrong:**
97
- - 用 technical reasoning push back
98
- - 展示证明其有效的 code/tests
99
- - 请求 clarification
100
-
101
- Template 见:requesting-code-review/code-reviewer.md
1
+ ---
2
+ name: requesting-code-review
3
+ description: 在完成任务、实现 major features,或 merge 前验证 work 是否满足 requirements 时使用
4
+ ---
5
+
6
+ # Requesting Code Review
7
+
8
+ Dispatch code reviewer subagent,在问题级联前捕获 issue。Reviewer 获得精确 crafted 的 evaluation context — 绝不是你的 session history。这使 reviewer 聚焦 work product,而非你的 thought process,并保留你自己的 context 以继续工作。
9
+
10
+ **Core principle:** Review early, review often.
11
+
12
+ ## When to Request Review
13
+
14
+ **Mandatory:**
15
+ - subagent-driven development 中每个 task 之后
16
+ - 完成 major feature 之后
17
+ - merge 到 main 之前
18
+
19
+ **Optional but valuable:**
20
+ - 卡住时(fresh perspective)
21
+ - refactoring 前(baseline check)
22
+ - 修复 complex bug 之后
23
+
24
+ ## How to Request
25
+
26
+ **1. Get git SHAs:**
27
+ ```bash
28
+ BASE_SHA=$(git rev-parse HEAD~1) # or origin/main
29
+ HEAD_SHA=$(git rev-parse HEAD)
30
+ ```
31
+
32
+ **2. Use the code reviewer template**(本 skill 目录下的 `code-reviewer.md`):
33
+
34
+ **Placeholders:**
35
+ - `{DESCRIPTION}` — 简要 summary of what you built
36
+ - `{PLAN_OR_REQUIREMENTS}` — 它应做什么(contract、exec plan 或 requirements)
37
+ - `{BASE_SHA}` — Starting commit
38
+ - `{HEAD_SHA}` — Ending commit
39
+
40
+ **3. Act on feedback:**
41
+ - Critical issues 立即修复
42
+ - Important issues 在继续前修复
43
+ - Minor issues 稍后处理
44
+ - Reviewer 有误时 push back(附 reasoning)
45
+
46
+ ## Example
47
+
48
+ ```
49
+ [Just completed Task 2: Add verification function]
50
+
51
+ You: Let me request code review before proceeding.
52
+
53
+ BASE_SHA=$(git log --oneline | grep "Task 1" | head -1 | awk '{print $1}')
54
+ HEAD_SHA=$(git rev-parse HEAD)
55
+
56
+ [Dispatch code reviewer subagent]
57
+ DESCRIPTION: Added verifyIndex() and repairIndex() with 4 issue types
58
+ PLAN_OR_REQUIREMENTS: Task 2 from docs/exec-plans/active/deployment-plan.md
59
+ BASE_SHA: a7981ec
60
+ HEAD_SHA: 3df7661
61
+
62
+ [Subagent returns]:
63
+ Strengths: Clean architecture, real tests
64
+ Issues:
65
+ Important: Missing progress indicators
66
+ Minor: Magic number (100) for reporting interval
67
+ Assessment: Ready to proceed
68
+
69
+ You: [Fix progress indicators]
70
+ [Continue to Task 3]
71
+ ```
72
+
73
+ ## Integration with Harness Workflow
74
+
75
+ **每个 work chunk 之后(Plan → Contract → Implement → Verify → Handoff):**
76
+ - Implement 之后、Verify 之前 review
77
+ - 在问题 compound 前捕获
78
+ - 进入 next task 前修复
79
+
80
+ **Before merge / Handoff:**
81
+ - 宣称 complete 前 review
82
+ - 对照 contract acceptance criteria 验证
83
+
84
+ **Ad-Hoc Development:**
85
+ - merge 前 review
86
+ - 卡住时 review
87
+
88
+ ## Red Flags
89
+
90
+ **Never:**
91
+ - 因 "it's simple" 跳过 review
92
+ - 忽略 Critical issues
93
+ - 带着未修复的 Important issues 继续
94
+ - 与 valid technical feedback 争辩
95
+
96
+ **If reviewer wrong:**
97
+ - 用 technical reasoning push back
98
+ - 展示证明其有效的 code/tests
99
+ - 请求 clarification
100
+
101
+ Template 见:requesting-code-review/code-reviewer.md
@@ -1,168 +1,168 @@
1
- # Code Reviewer Prompt Template
2
-
3
- Dispatch code reviewer subagent 时使用本 template。
4
-
5
- **Purpose:** 在 work cascade 成更多工作之前,对照 requirements 与 code quality standards review completed work。
6
-
7
- ```
8
- Task tool (general-purpose):
9
- description: "Review code changes"
10
- prompt: |
11
- You are a Senior Code Reviewer with expertise in software architecture,
12
- design patterns, and best practices. Your job is to review completed work
13
- against its plan or requirements and identify issues before they cascade.
14
-
15
- ## What Was Implemented
16
-
17
- {DESCRIPTION}
18
-
19
- ## Requirements / Plan
20
-
21
- {PLAN_OR_REQUIREMENTS}
22
-
23
- ## Git Range to Review
24
-
25
- **Base:** {BASE_SHA}
26
- **Head:** {HEAD_SHA}
27
-
28
- ```bash
29
- git diff --stat {BASE_SHA}..{HEAD_SHA}
30
- git diff {BASE_SHA}..{HEAD_SHA}
31
- ```
32
-
33
- ## What to Check
34
-
35
- **Plan alignment:**
36
- - Does the implementation match the plan / requirements?
37
- - Are deviations justified improvements, or problematic departures?
38
- - Is all planned functionality present?
39
-
40
- **Code quality:**
41
- - Clean separation of concerns?
42
- - Proper error handling?
43
- - Type safety where applicable?
44
- - DRY without premature abstraction?
45
- - Edge cases handled?
46
-
47
- **Architecture:**
48
- - Sound design decisions?
49
- - Reasonable scalability and performance?
50
- - Security concerns?
51
- - Integrates cleanly with surrounding code?
52
-
53
- **Testing:**
54
- - Tests verify real behavior, not mocks?
55
- - Edge cases covered?
56
- - Integration tests where they matter?
57
- - All tests passing?
58
-
59
- **Production readiness:**
60
- - Migration strategy if schema changed?
61
- - Backward compatibility considered?
62
- - Documentation complete?
63
- - No obvious bugs?
64
-
65
- ## Calibration
66
-
67
- Categorize issues by actual severity. Not everything is Critical.
68
- Acknowledge what was done well before listing issues — accurate praise
69
- helps the implementer trust the rest of the feedback.
70
-
71
- If you find significant deviations from the plan, flag them specifically
72
- so the implementer can confirm whether the deviation was intentional.
73
- If you find issues with the plan itself rather than the implementation,
74
- say so.
75
-
76
- ## Output Format
77
-
78
- ### Strengths
79
- [What's well done? Be specific.]
80
-
81
- ### Issues
82
-
83
- #### Critical (Must Fix)
84
- [Bugs, security issues, data loss risks, broken functionality]
85
-
86
- #### Important (Should Fix)
87
- [Architecture problems, missing features, poor error handling, test gaps]
88
-
89
- #### Minor (Nice to Have)
90
- [Code style, optimization opportunities, documentation polish]
91
-
92
- For each issue:
93
- - File:line reference
94
- - What's wrong
95
- - Why it matters
96
- - How to fix (if not obvious)
97
-
98
- ### Recommendations
99
- [Improvements for code quality, architecture, or process]
100
-
101
- ### Assessment
102
-
103
- **Ready to merge?** [Yes | No | With fixes]
104
-
105
- **Reasoning:** [1-2 sentence technical assessment]
106
-
107
- ## Critical Rules
108
-
109
- **DO:**
110
- - Categorize by actual severity
111
- - Be specific (file:line, not vague)
112
- - Explain WHY each issue matters
113
- - Acknowledge strengths
114
- - Give a clear verdict
115
-
116
- **DON'T:**
117
- - Say "looks good" without checking
118
- - Mark nitpicks as Critical
119
- - Give feedback on code you didn't actually read
120
- - Be vague ("improve error handling")
121
- - Avoid giving a clear verdict
122
- ```
123
-
124
- **Placeholders:**
125
- - `{DESCRIPTION}` — brief summary of what was built
126
- - `{PLAN_OR_REQUIREMENTS}` — 它应做什么(plan file path、task text 或 requirements)
127
- - `{BASE_SHA}` — starting commit
128
- - `{HEAD_SHA}` — ending commit
129
-
130
- **Reviewer returns:** Strengths、Issues (Critical / Important / Minor)、Recommendations、Assessment
131
-
132
- ## Example Output
133
-
134
- ```
135
- ### Strengths
136
- - Clean database schema with proper migrations (db.ts:15-42)
137
- - Comprehensive test coverage (18 tests, all edge cases)
138
- - Good error handling with fallbacks (summarizer.ts:85-92)
139
-
140
- ### Issues
141
-
142
- #### Important
143
- 1. **Missing help text in CLI wrapper**
144
- - File: index-conversations:1-31
145
- - Issue: No --help flag, users won't discover --concurrency
146
- - Fix: Add --help case with usage examples
147
-
148
- 2. **Date validation missing**
149
- - File: search.ts:25-27
150
- - Issue: Invalid dates silently return no results
151
- - Fix: Validate ISO format, throw error with example
152
-
153
- #### Minor
154
- 1. **Progress indicators**
155
- - File: indexer.ts:130
156
- - Issue: No "X of Y" counter for long operations
157
- - Impact: Users don't know how long to wait
158
-
159
- ### Recommendations
160
- - Add progress reporting for user experience
161
- - Consider config file for excluded projects (portability)
162
-
163
- ### Assessment
164
-
165
- **Ready to merge: With fixes**
166
-
167
- **Reasoning:** Core implementation is solid with good architecture and tests. Important issues (help text, date validation) are easily fixed and don't affect core functionality.
168
- ```
1
+ # Code Reviewer Prompt Template
2
+
3
+ Dispatch code reviewer subagent 时使用本 template。
4
+
5
+ **Purpose:** 在 work cascade 成更多工作之前,对照 requirements 与 code quality standards review completed work。
6
+
7
+ ```
8
+ Task tool (general-purpose):
9
+ description: "Review code changes"
10
+ prompt: |
11
+ You are a Senior Code Reviewer with expertise in software architecture,
12
+ design patterns, and best practices. Your job is to review completed work
13
+ against its plan or requirements and identify issues before they cascade.
14
+
15
+ ## What Was Implemented
16
+
17
+ {DESCRIPTION}
18
+
19
+ ## Requirements / Plan
20
+
21
+ {PLAN_OR_REQUIREMENTS}
22
+
23
+ ## Git Range to Review
24
+
25
+ **Base:** {BASE_SHA}
26
+ **Head:** {HEAD_SHA}
27
+
28
+ ```bash
29
+ git diff --stat {BASE_SHA}..{HEAD_SHA}
30
+ git diff {BASE_SHA}..{HEAD_SHA}
31
+ ```
32
+
33
+ ## What to Check
34
+
35
+ **Plan alignment:**
36
+ - Does the implementation match the plan / requirements?
37
+ - Are deviations justified improvements, or problematic departures?
38
+ - Is all planned functionality present?
39
+
40
+ **Code quality:**
41
+ - Clean separation of concerns?
42
+ - Proper error handling?
43
+ - Type safety where applicable?
44
+ - DRY without premature abstraction?
45
+ - Edge cases handled?
46
+
47
+ **Architecture:**
48
+ - Sound design decisions?
49
+ - Reasonable scalability and performance?
50
+ - Security concerns?
51
+ - Integrates cleanly with surrounding code?
52
+
53
+ **Testing:**
54
+ - Tests verify real behavior, not mocks?
55
+ - Edge cases covered?
56
+ - Integration tests where they matter?
57
+ - All tests passing?
58
+
59
+ **Production readiness:**
60
+ - Migration strategy if schema changed?
61
+ - Backward compatibility considered?
62
+ - Documentation complete?
63
+ - No obvious bugs?
64
+
65
+ ## Calibration
66
+
67
+ Categorize issues by actual severity. Not everything is Critical.
68
+ Acknowledge what was done well before listing issues — accurate praise
69
+ helps the implementer trust the rest of the feedback.
70
+
71
+ If you find significant deviations from the plan, flag them specifically
72
+ so the implementer can confirm whether the deviation was intentional.
73
+ If you find issues with the plan itself rather than the implementation,
74
+ say so.
75
+
76
+ ## Output Format
77
+
78
+ ### Strengths
79
+ [What's well done? Be specific.]
80
+
81
+ ### Issues
82
+
83
+ #### Critical (Must Fix)
84
+ [Bugs, security issues, data loss risks, broken functionality]
85
+
86
+ #### Important (Should Fix)
87
+ [Architecture problems, missing features, poor error handling, test gaps]
88
+
89
+ #### Minor (Nice to Have)
90
+ [Code style, optimization opportunities, documentation polish]
91
+
92
+ For each issue:
93
+ - File:line reference
94
+ - What's wrong
95
+ - Why it matters
96
+ - How to fix (if not obvious)
97
+
98
+ ### Recommendations
99
+ [Improvements for code quality, architecture, or process]
100
+
101
+ ### Assessment
102
+
103
+ **Ready to merge?** [Yes | No | With fixes]
104
+
105
+ **Reasoning:** [1-2 sentence technical assessment]
106
+
107
+ ## Critical Rules
108
+
109
+ **DO:**
110
+ - Categorize by actual severity
111
+ - Be specific (file:line, not vague)
112
+ - Explain WHY each issue matters
113
+ - Acknowledge strengths
114
+ - Give a clear verdict
115
+
116
+ **DON'T:**
117
+ - Say "looks good" without checking
118
+ - Mark nitpicks as Critical
119
+ - Give feedback on code you didn't actually read
120
+ - Be vague ("improve error handling")
121
+ - Avoid giving a clear verdict
122
+ ```
123
+
124
+ **Placeholders:**
125
+ - `{DESCRIPTION}` — brief summary of what was built
126
+ - `{PLAN_OR_REQUIREMENTS}` — 它应做什么(plan file path、task text 或 requirements)
127
+ - `{BASE_SHA}` — starting commit
128
+ - `{HEAD_SHA}` — ending commit
129
+
130
+ **Reviewer returns:** Strengths、Issues (Critical / Important / Minor)、Recommendations、Assessment
131
+
132
+ ## Example Output
133
+
134
+ ```
135
+ ### Strengths
136
+ - Clean database schema with proper migrations (db.ts:15-42)
137
+ - Comprehensive test coverage (18 tests, all edge cases)
138
+ - Good error handling with fallbacks (summarizer.ts:85-92)
139
+
140
+ ### Issues
141
+
142
+ #### Important
143
+ 1. **Missing help text in CLI wrapper**
144
+ - File: index-conversations:1-31
145
+ - Issue: No --help flag, users won't discover --concurrency
146
+ - Fix: Add --help case with usage examples
147
+
148
+ 2. **Date validation missing**
149
+ - File: search.ts:25-27
150
+ - Issue: Invalid dates silently return no results
151
+ - Fix: Validate ISO format, throw error with example
152
+
153
+ #### Minor
154
+ 1. **Progress indicators**
155
+ - File: indexer.ts:130
156
+ - Issue: No "X of Y" counter for long operations
157
+ - Impact: Users don't know how long to wait
158
+
159
+ ### Recommendations
160
+ - Add progress reporting for user experience
161
+ - Consider config file for excluded projects (portability)
162
+
163
+ ### Assessment
164
+
165
+ **Ready to merge: With fixes**
166
+
167
+ **Reasoning:** Core implementation is solid with good architecture and tests. Important issues (help text, date validation) are easily fixed and don't affect core functionality.
168
+ ```