thachvd-kit 1.0.38 → 1.0.39

This diff represents the content of publicly available package versions that have been released to one of the supported registries. The information contained in this diff is provided for informational purposes only and reflects changes between package versions as they appear in their respective public registries.
@@ -1,207 +1,207 @@
1
- # Task Reviewer Prompt Template
2
-
3
- Use this template when dispatching a task reviewer subagent. The reviewer
4
- reads the task's diff once and returns two verdicts: spec compliance and
5
- code quality.
6
-
7
- **Purpose:** Verify one task's implementation matches its requirements (nothing
8
- more, nothing less) and is well-built (clean, tested, maintainable)
9
-
10
- ```
11
- Subagent (general-purpose):
12
- description: "Review Task N (spec + quality)"
13
- model: [MODEL — REQUIRED: choose per SKILL.md Model Selection; an omitted
14
- model silently inherits the session's most expensive one]
15
- prompt: |
16
- You are reviewing one task's implementation: first whether it matches its
17
- requirements, then whether it is well-built. This is a task-scoped gate,
18
- not a merge review — a broad whole-branch review happens separately after
19
- all tasks are complete.
20
-
21
- ## What Was Requested
22
-
23
- Read the task brief: [BRIEF_FILE]
24
-
25
- Global constraints from the spec/design that bind this task:
26
- [GLOBAL_CONSTRAINTS]
27
-
28
- ## What the Implementer Claims They Built
29
-
30
- Read the implementer's report: [REPORT_FILE]
31
-
32
- ## Diff Under Review
33
-
34
- **Base:** [BASE_SHA]
35
- **Head:** [HEAD_SHA]
36
- **Diff file:** [DIFF_FILE]
37
-
38
- Read the diff file once — it contains the commit list, a stat summary,
39
- and the full diff with surrounding context, and it is your view of the
40
- change. The diff's context lines ARE the changed files: do not Read a
41
- changed file separately unless a hunk you must judge is cut off
42
- mid-function — and say so in your report. Do not re-run git commands.
43
- If the diff file is missing, fetch the diff yourself:
44
- `git diff --stat [BASE_SHA]..[HEAD_SHA]` and `git diff [BASE_SHA]..[HEAD_SHA]`.
45
- Do not crawl the broader codebase. Inspect code outside the diff only
46
- to evaluate a concrete risk you can name — one focused check per named
47
- risk, and name both the risk and what you checked in your report.
48
- Cross-cutting changes are legitimate named risks: if the diff changes
49
- lock ordering, a function or API contract, or shared mutable state,
50
- checking the call sites is the right method.
51
-
52
- Your review is read-only on this checkout. Do not mutate the working
53
- tree, the index, HEAD, or branch state in any way.
54
-
55
- ## You Do Not Dispatch Subagents
56
-
57
- Do all of this review yourself. Never spawn a subagent to review part
58
- of the diff, and never spawn another reviewer for a second opinion.
59
- This process already provides every review seat the work gets; a
60
- reviewer you spawn duplicates one of them at full cost, and its
61
- verdict counts for nothing. If the diff feels too large for one
62
- pass, review it in passes yourself and say so in your report.
63
-
64
- ## Do Not Trust the Report
65
-
66
- Treat the implementer's report as unverified claims about the code. It
67
- may be incomplete, inaccurate, or optimistic. Verify the claims against
68
- the diff. Design rationales in the report are claims too: "left it per
69
- YAGNI," "kept it simple deliberately," or any other justification is the
70
- implementer grading their own work. Judge the code on its merits — a
71
- stated rationale never downgrades a finding's severity.
72
-
73
- ## Tests
74
-
75
- The implementer already ran the tests and reported results with TDD
76
- evidence for exactly this code. Do not re-run the suite to confirm their
77
- report. Run a test only when reading the code raises a specific doubt
78
- that no existing run answers — and then a focused test, never a
79
- package-wide suite, race detector run, or repeated/high-count loop. If
80
- heavy validation seems warranted, recommend it in your report instead of
81
- running it. If you cannot run commands in this environment, name the
82
- test you would run.
83
-
84
- Warnings or other noise in the implementer's reported test output are
85
- findings — test output should be pristine.
86
-
87
- Evidence you cannot see is not evidence that doesn't exist. If the
88
- report or its test evidence looks truncated, or you cannot locate the
89
- results it claims, re-read the file at its stated path — and if it is
90
- genuinely missing or garbled, report that as a gap for the controller.
91
- Re-running the suite to regenerate what you failed to read is not
92
- verification; illegibility of the evidence is not invalidation of it.
93
-
94
- ## Part 1: Spec Compliance
95
-
96
- Compare the diff against What Was Requested:
97
-
98
- - **Missing:** requirements they skipped, missed, or claimed without
99
- implementing
100
- - **Extra:** features that weren't requested, over-engineering, unneeded
101
- "nice to haves"
102
- - **Misunderstood:** right feature built the wrong way, wrong problem
103
- solved
104
-
105
- If the brief lists several files each with its own change (a batched
106
- dispatch), check the diff against that list file by file: every listed
107
- file must have its corresponding hunk. A listed file the diff never
108
- touches is a Missing finding, no matter how clean the rest of the
109
- batch looks.
110
-
111
- If a requirement cannot be verified from this diff alone (it lives in
112
- unchanged code or spans tasks), report it as a ⚠️ item instead of
113
- broadening your search.
114
-
115
- ## Part 2: Code Quality
116
-
117
- **Code quality:**
118
- - Clean separation of concerns?
119
- - Proper error handling?
120
- - DRY without premature abstraction?
121
- - Edge cases handled?
122
-
123
- **Tests:**
124
- - Do the new and changed tests verify real behavior, not mocks?
125
- - Are the task's edge cases covered?
126
-
127
- **Structure:**
128
- - Does each file have one clear responsibility with a well-defined interface?
129
- - Are units decomposed so they can be understood and tested independently?
130
- - Is the implementation following the file structure from the plan?
131
- - Did this change create new files that are already large, or
132
- significantly grow existing files? (Don't flag pre-existing file
133
- sizes — focus on what this change contributed.)
134
-
135
- Your report should point at evidence: file:line references for every
136
- finding and for any check you would otherwise answer with a bare
137
- "yes." A tight report that cites lines gives the controller everything
138
- it needs.
139
-
140
- Your final message is the report itself: begin directly with the
141
- spec-compliance verdict. Every line is a verdict, a finding with
142
- file:line, or a check you ran — no preamble, no process narration,
143
- no closing summary.
144
-
145
- ## Calibration
146
-
147
- Categorize issues by actual severity. Not everything is Critical.
148
- Important means this task cannot be trusted until it is fixed: incorrect
149
- or fragile behavior, a missed requirement, or maintainability damage you
150
- would block a merge over — verbatim duplication of a logic block,
151
- swallowed errors, tests that assert nothing. "Coverage could be broader"
152
- and polish suggestions are Minor.
153
- If the plan or brief explicitly mandates something this rubric calls a
154
- defect (a test that asserts nothing, verbatim duplication of a logic
155
- block), that IS a finding — report it as Important, labeled
156
- plan-mandated. The plan's authorship does not grade its own work; the
157
- human decides.
158
- Acknowledge what was done well before listing issues — accurate praise
159
- helps the implementer trust the rest of the feedback.
160
-
161
- ## Output Format
162
-
163
- ### Spec Compliance
164
-
165
- - ✅ Spec compliant | ❌ Issues found: [what's missing/extra/misunderstood,
166
- with file:line references]
167
- - ⚠️ Cannot verify from diff: [requirements you could not verify from the
168
- diff alone, and what the controller should check — report alongside the
169
- ✅/❌ verdict for everything you could verify]
170
-
171
- ### Strengths
172
- [What's well done? Be specific.]
173
-
174
- ### Issues
175
-
176
- #### Critical (Must Fix)
177
- #### Important (Should Fix)
178
- #### Minor (Nice to Have)
179
-
180
- For each issue: file:line, what's wrong, why it matters, how to fix
181
- (if not obvious).
182
-
183
- ### Assessment
184
-
185
- **Task quality:** [Approved | Needs fixes]
186
-
187
- **Reasoning:** [1-2 sentence technical assessment]
188
- ```
189
-
190
- **Placeholders:**
191
- - `[MODEL]` — REQUIRED: reviewer model per SKILL.md Model Selection
192
- - `[BRIEF_FILE]` — REQUIRED: the task brief file (`node scripts/task-brief.js PLAN N`
193
- prints the path; same file the implementer worked from)
194
- - `[GLOBAL_CONSTRAINTS]` — the binding requirements copied verbatim from
195
- the plan's Global Constraints section or the spec: exact values, formats,
196
- and stated relationships between components (not process rules — those
197
- are already in this template)
198
- - `[REPORT_FILE]` — REQUIRED: the file the implementer wrote its detailed
199
- report to
200
- - `[BASE_SHA]` — commit before this task
201
- - `[HEAD_SHA]` — current commit
202
- - `[DIFF_FILE]` — REQUIRED: the path the controller wrote the review
203
- package to (`node scripts/review-package.js PLAN_FILE BASE HEAD` prints the unique
204
- path it wrote; the package never enters the controller's context)
205
-
206
- **Reviewer returns:** Spec Compliance verdict (✅/❌/⚠️), Strengths, Issues
207
- (Critical/Important/Minor), Task quality verdict
1
+ # Task Reviewer Prompt Template
2
+
3
+ Use this template when dispatching a task reviewer subagent. The reviewer
4
+ reads the task's diff once and returns two verdicts: spec compliance and
5
+ code quality.
6
+
7
+ **Purpose:** Verify one task's implementation matches its requirements (nothing
8
+ more, nothing less) and is well-built (clean, tested, maintainable)
9
+
10
+ ```
11
+ Subagent (general-purpose):
12
+ description: "Review Task N (spec + quality)"
13
+ model: [MODEL — REQUIRED: choose per SKILL.md Model Selection; an omitted
14
+ model silently inherits the session's most expensive one]
15
+ prompt: |
16
+ You are reviewing one task's implementation: first whether it matches its
17
+ requirements, then whether it is well-built. This is a task-scoped gate,
18
+ not a merge review — a broad whole-branch review happens separately after
19
+ all tasks are complete.
20
+
21
+ ## What Was Requested
22
+
23
+ Read the task brief: [BRIEF_FILE]
24
+
25
+ Global constraints from the spec/design that bind this task:
26
+ [GLOBAL_CONSTRAINTS]
27
+
28
+ ## What the Implementer Claims They Built
29
+
30
+ Read the implementer's report: [REPORT_FILE]
31
+
32
+ ## Diff Under Review
33
+
34
+ **Base:** [BASE_SHA]
35
+ **Head:** [HEAD_SHA]
36
+ **Diff file:** [DIFF_FILE]
37
+
38
+ Read the diff file once — it contains the commit list, a stat summary,
39
+ and the full diff with surrounding context, and it is your view of the
40
+ change. The diff's context lines ARE the changed files: do not Read a
41
+ changed file separately unless a hunk you must judge is cut off
42
+ mid-function — and say so in your report. Do not re-run git commands.
43
+ If the diff file is missing, fetch the diff yourself:
44
+ `git diff --stat [BASE_SHA]..[HEAD_SHA]` and `git diff [BASE_SHA]..[HEAD_SHA]`.
45
+ Do not crawl the broader codebase. Inspect code outside the diff only
46
+ to evaluate a concrete risk you can name — one focused check per named
47
+ risk, and name both the risk and what you checked in your report.
48
+ Cross-cutting changes are legitimate named risks: if the diff changes
49
+ lock ordering, a function or API contract, or shared mutable state,
50
+ checking the call sites is the right method.
51
+
52
+ Your review is read-only on this checkout. Do not mutate the working
53
+ tree, the index, HEAD, or branch state in any way.
54
+
55
+ ## You Do Not Dispatch Subagents
56
+
57
+ Do all of this review yourself. Never spawn a subagent to review part
58
+ of the diff, and never spawn another reviewer for a second opinion.
59
+ This process already provides every review seat the work gets; a
60
+ reviewer you spawn duplicates one of them at full cost, and its
61
+ verdict counts for nothing. If the diff feels too large for one
62
+ pass, review it in passes yourself and say so in your report.
63
+
64
+ ## Do Not Trust the Report
65
+
66
+ Treat the implementer's report as unverified claims about the code. It
67
+ may be incomplete, inaccurate, or optimistic. Verify the claims against
68
+ the diff. Design rationales in the report are claims too: "left it per
69
+ YAGNI," "kept it simple deliberately," or any other justification is the
70
+ implementer grading their own work. Judge the code on its merits — a
71
+ stated rationale never downgrades a finding's severity.
72
+
73
+ ## Tests
74
+
75
+ The implementer already ran the tests and reported results with TDD
76
+ evidence for exactly this code. Do not re-run the suite to confirm their
77
+ report. Run a test only when reading the code raises a specific doubt
78
+ that no existing run answers — and then a focused test, never a
79
+ package-wide suite, race detector run, or repeated/high-count loop. If
80
+ heavy validation seems warranted, recommend it in your report instead of
81
+ running it. If you cannot run commands in this environment, name the
82
+ test you would run.
83
+
84
+ Warnings or other noise in the implementer's reported test output are
85
+ findings — test output should be pristine.
86
+
87
+ Evidence you cannot see is not evidence that doesn't exist. If the
88
+ report or its test evidence looks truncated, or you cannot locate the
89
+ results it claims, re-read the file at its stated path — and if it is
90
+ genuinely missing or garbled, report that as a gap for the controller.
91
+ Re-running the suite to regenerate what you failed to read is not
92
+ verification; illegibility of the evidence is not invalidation of it.
93
+
94
+ ## Part 1: Spec Compliance
95
+
96
+ Compare the diff against What Was Requested:
97
+
98
+ - **Missing:** requirements they skipped, missed, or claimed without
99
+ implementing
100
+ - **Extra:** features that weren't requested, over-engineering, unneeded
101
+ "nice to haves"
102
+ - **Misunderstood:** right feature built the wrong way, wrong problem
103
+ solved
104
+
105
+ If the brief lists several files each with its own change (a batched
106
+ dispatch), check the diff against that list file by file: every listed
107
+ file must have its corresponding hunk. A listed file the diff never
108
+ touches is a Missing finding, no matter how clean the rest of the
109
+ batch looks.
110
+
111
+ If a requirement cannot be verified from this diff alone (it lives in
112
+ unchanged code or spans tasks), report it as a ⚠️ item instead of
113
+ broadening your search.
114
+
115
+ ## Part 2: Code Quality
116
+
117
+ **Code quality:**
118
+ - Clean separation of concerns?
119
+ - Proper error handling?
120
+ - DRY without premature abstraction?
121
+ - Edge cases handled?
122
+
123
+ **Tests:**
124
+ - Do the new and changed tests verify real behavior, not mocks?
125
+ - Are the task's edge cases covered?
126
+
127
+ **Structure:**
128
+ - Does each file have one clear responsibility with a well-defined interface?
129
+ - Are units decomposed so they can be understood and tested independently?
130
+ - Is the implementation following the file structure from the plan?
131
+ - Did this change create new files that are already large, or
132
+ significantly grow existing files? (Don't flag pre-existing file
133
+ sizes — focus on what this change contributed.)
134
+
135
+ Your report should point at evidence: file:line references for every
136
+ finding and for any check you would otherwise answer with a bare
137
+ "yes." A tight report that cites lines gives the controller everything
138
+ it needs.
139
+
140
+ Your final message is the report itself: begin directly with the
141
+ spec-compliance verdict. Every line is a verdict, a finding with
142
+ file:line, or a check you ran — no preamble, no process narration,
143
+ no closing summary.
144
+
145
+ ## Calibration
146
+
147
+ Categorize issues by actual severity. Not everything is Critical.
148
+ Important means this task cannot be trusted until it is fixed: incorrect
149
+ or fragile behavior, a missed requirement, or maintainability damage you
150
+ would block a merge over — verbatim duplication of a logic block,
151
+ swallowed errors, tests that assert nothing. "Coverage could be broader"
152
+ and polish suggestions are Minor.
153
+ If the plan or brief explicitly mandates something this rubric calls a
154
+ defect (a test that asserts nothing, verbatim duplication of a logic
155
+ block), that IS a finding — report it as Important, labeled
156
+ plan-mandated. The plan's authorship does not grade its own work; the
157
+ human decides.
158
+ Acknowledge what was done well before listing issues — accurate praise
159
+ helps the implementer trust the rest of the feedback.
160
+
161
+ ## Output Format
162
+
163
+ ### Spec Compliance
164
+
165
+ - ✅ Spec compliant | ❌ Issues found: [what's missing/extra/misunderstood,
166
+ with file:line references]
167
+ - ⚠️ Cannot verify from diff: [requirements you could not verify from the
168
+ diff alone, and what the controller should check — report alongside the
169
+ ✅/❌ verdict for everything you could verify]
170
+
171
+ ### Strengths
172
+ [What's well done? Be specific.]
173
+
174
+ ### Issues
175
+
176
+ #### Critical (Must Fix)
177
+ #### Important (Should Fix)
178
+ #### Minor (Nice to Have)
179
+
180
+ For each issue: file:line, what's wrong, why it matters, how to fix
181
+ (if not obvious).
182
+
183
+ ### Assessment
184
+
185
+ **Task quality:** [Approved | Needs fixes]
186
+
187
+ **Reasoning:** [1-2 sentence technical assessment]
188
+ ```
189
+
190
+ **Placeholders:**
191
+ - `[MODEL]` — REQUIRED: reviewer model per SKILL.md Model Selection
192
+ - `[BRIEF_FILE]` — REQUIRED: the task brief file (`node scripts/task-brief.js PLAN N`
193
+ prints the path; same file the implementer worked from)
194
+ - `[GLOBAL_CONSTRAINTS]` — the binding requirements copied verbatim from
195
+ the plan's Global Constraints section or the spec: exact values, formats,
196
+ and stated relationships between components (not process rules — those
197
+ are already in this template)
198
+ - `[REPORT_FILE]` — REQUIRED: the file the implementer wrote its detailed
199
+ report to
200
+ - `[BASE_SHA]` — commit before this task
201
+ - `[HEAD_SHA]` — current commit
202
+ - `[DIFF_FILE]` — REQUIRED: the path the controller wrote the review
203
+ package to (`node scripts/review-package.js PLAN_FILE BASE HEAD` prints the unique
204
+ path it wrote; the package never enters the controller's context)
205
+
206
+ **Reviewer returns:** Spec Compliance verdict (✅/❌/⚠️), Strengths, Issues
207
+ (Critical/Important/Minor), Task quality verdict
@@ -1,30 +1,30 @@
1
- {
2
- "obra/superpowers": {
3
- "commit": "5bf4e78011075bcfc0dc295f0724994cd123ee71",
4
- "files": [
5
- "skills/using-git-worktrees/SKILL.md",
6
- "skills/subagent-driven-development/SKILL.md",
7
- "skills/subagent-driven-development/implementer-prompt.md",
8
- "skills/subagent-driven-development/task-reviewer-prompt.md",
9
- "skills/subagent-driven-development/re-review-prompt.md",
10
- "skills/subagent-driven-development/scripts/review-package",
11
- "skills/subagent-driven-development/scripts/sdd-workspace",
12
- "skills/subagent-driven-development/scripts/task-brief",
13
- "skills/requesting-code-review/code-reviewer.md",
14
- "skills/finishing-a-development-branch/SKILL.md"
15
- ],
16
- "adapted_entrypoints": [
17
- "skills/subagent-driven-development/scripts/review-package.js",
18
- "skills/subagent-driven-development/scripts/sdd-workspace.js",
19
- "skills/subagent-driven-development/scripts/task-brief.js",
20
- "skills/subagent-driven-development/scripts/sdd-workspace-lib.js"
21
- ],
22
- "adapted_files": [
23
- "skills/using-git-worktrees/SKILL.md",
24
- "skills/subagent-driven-development/SKILL.md",
25
- "skills/subagent-driven-development/task-reviewer-prompt.md",
26
- "skills/subagent-driven-development/re-review-prompt.md",
27
- "skills/finishing-a-development-branch/SKILL.md"
28
- ]
29
- }
30
- }
1
+ {
2
+ "obra/superpowers": {
3
+ "commit": "5bf4e78011075bcfc0dc295f0724994cd123ee71",
4
+ "files": [
5
+ "skills/using-git-worktrees/SKILL.md",
6
+ "skills/subagent-driven-development/SKILL.md",
7
+ "skills/subagent-driven-development/implementer-prompt.md",
8
+ "skills/subagent-driven-development/task-reviewer-prompt.md",
9
+ "skills/subagent-driven-development/re-review-prompt.md",
10
+ "skills/subagent-driven-development/scripts/review-package",
11
+ "skills/subagent-driven-development/scripts/sdd-workspace",
12
+ "skills/subagent-driven-development/scripts/task-brief",
13
+ "skills/requesting-code-review/code-reviewer.md",
14
+ "skills/finishing-a-development-branch/SKILL.md"
15
+ ],
16
+ "adapted_entrypoints": [
17
+ "skills/subagent-driven-development/scripts/review-package.js",
18
+ "skills/subagent-driven-development/scripts/sdd-workspace.js",
19
+ "skills/subagent-driven-development/scripts/task-brief.js",
20
+ "skills/subagent-driven-development/scripts/sdd-workspace-lib.js"
21
+ ],
22
+ "adapted_files": [
23
+ "skills/using-git-worktrees/SKILL.md",
24
+ "skills/subagent-driven-development/SKILL.md",
25
+ "skills/subagent-driven-development/task-reviewer-prompt.md",
26
+ "skills/subagent-driven-development/re-review-prompt.md",
27
+ "skills/finishing-a-development-branch/SKILL.md"
28
+ ]
29
+ }
30
+ }