@webpresso/plugin-claude 0.0.10 → 0.0.11
This diff represents the content of publicly available package versions that have been released to one of the supported registries. The information contained in this diff is provided for informational purposes only and reflects changes between package versions as they appear in their respective public registries.
- package/.claude-plugin/marketplace.json +9 -2
- package/.claude-plugin/plugin.json +1 -1
- package/bin/wp +52 -58
- package/hooks/hooks.json +102 -0
- package/package.json +3 -1
- package/plugin-skill-ownership.json +14 -14
- package/plugins/typescript-lsp/.claude-plugin/plugin.json +25 -0
- package/skills/autopilot/SKILL.md +1 -1
- package/skills/claude/SKILL.md +24 -166
- package/skills/codex/SKILL.md +25 -123
- package/skills/grok/SKILL.md +30 -18
- package/skills/hooks-doctor/SKILL.md +28 -18
- package/skills/opencode-go/SKILL.md +35 -85
- package/skills/plan-refine/SKILL.md +1 -0
- package/skills/plan-refine/references/full-methodology.md +18 -18
- package/skills/ralplan/SKILL.md +1 -0
- package/skills/tech-debt/SKILL.md +6 -0
- package/skills/testing-philosophy/references/full-testing-philosophy.md +8 -12
- package/skills/tooling-friction/SKILL.md +38 -16
- package/skills/ultragoal/SKILL.md +10 -9
- package/skills/verify/SKILL.md +57 -23
package/skills/codex/SKILL.md
CHANGED
|
@@ -1,142 +1,44 @@
|
|
|
1
1
|
---
|
|
2
2
|
name: codex
|
|
3
|
-
description: "Codex
|
|
3
|
+
description: "Codex outside-voice reviewer through Webpresso MCP."
|
|
4
4
|
license: MIT
|
|
5
5
|
---
|
|
6
6
|
|
|
7
|
-
# Codex outside
|
|
7
|
+
# Codex outside-voice review
|
|
8
8
|
|
|
9
|
-
Use
|
|
9
|
+
Use only for a requested content-bound plan or delivery review. Treat output as external advice
|
|
10
|
+
until independently verified.
|
|
10
11
|
|
|
11
|
-
##
|
|
12
|
+
## MCP-only contract
|
|
12
13
|
|
|
13
|
-
|
|
14
|
+
- Call `wp_review_gate` exactly once for the real plan version or delivery diff.
|
|
15
|
+
- Observe the same durable operation with `wp_review_gate_wait` when needed.
|
|
16
|
+
- Never invoke a provider command, the legacy review CLI, or a removed review-run surface.
|
|
17
|
+
- If the review MCP is unavailable, stop and report it unavailable. There is no CLI fallback.
|
|
18
|
+
|
|
19
|
+
## Advisory call
|
|
14
20
|
|
|
15
21
|
```jsonc
|
|
16
|
-
// wp_review_run MCP tool call
|
|
17
22
|
{
|
|
18
|
-
"
|
|
23
|
+
"project_id": "<project>",
|
|
24
|
+
"slug": "<blueprint-slug>",
|
|
25
|
+
"purpose": "delivery",
|
|
26
|
+
"base_ref": "<full-base-sha>",
|
|
27
|
+
"authority_ref": "<full-head-sha>",
|
|
19
28
|
"provider": "codex",
|
|
20
|
-
|
|
21
|
-
|
|
22
|
-
// review.
|
|
29
|
+
"repository_access": "none",
|
|
30
|
+
"mode": "advisory",
|
|
23
31
|
}
|
|
24
32
|
```
|
|
25
33
|
|
|
26
|
-
Use
|
|
27
|
-
|
|
28
|
-
For long-running reviews, prefer the async MCP path: call `wp_review_run` with
|
|
29
|
-
`"provider": "codex"` and `"background": true`, then poll `wp_review_wait`
|
|
30
|
-
with the returned `runId`. Do not solve MCP transport limits by inflating
|
|
31
|
-
synchronous review timeouts.
|
|
32
|
-
|
|
33
|
-
The bash blocks in this skill (below) are the **MCP-unavailable fallback only** — use them when the webpresso MCP server itself is not reachable in the current host, not as a provider-fallback mechanism.
|
|
34
|
-
|
|
35
|
-
## Single-shot budget (anti-stampede)
|
|
36
|
-
|
|
37
|
-
- Default: **one** review invocation per request.
|
|
38
|
-
- Do **not** fan out parallel multi-host review matrices unless the user set `review_budget`/`N` > 1.
|
|
39
|
-
- Prefer `wp_review_run` (MCP) or `wp review run` (CLI fallback) over spawn/wait agent loops for review.
|
|
40
|
-
- Keep prompts bounded; no whole-repo paste.
|
|
41
|
-
|
|
42
|
-
## MCP-unavailable fallback (raw CLI)
|
|
43
|
-
|
|
44
|
-
Everything from here down is the manual `wp review run` CLI path documented for hosts or sessions where the webpresso MCP server is not reachable. Prefer `wp_review_run` above whenever MCP is available.
|
|
45
|
-
|
|
46
|
-
## MCP-unavailable fallback: auth check
|
|
47
|
-
|
|
48
|
-
```bash
|
|
49
|
-
if ! codex login status >/dev/null 2>&1; then
|
|
50
|
-
echo "CODEX_AUTH=missing: run codex login before using the codex outside-voice skill"
|
|
51
|
-
exit 1
|
|
52
|
-
fi
|
|
53
|
-
echo "CODEX_AUTH=ok"
|
|
54
|
-
```
|
|
55
|
-
|
|
56
|
-
## MCP-unavailable fallback: portable prompt file
|
|
57
|
-
|
|
58
|
-
```bash
|
|
59
|
-
PROMPT_FILE=$(mktemp -t wp-codex-review.XXXXXX)
|
|
60
|
-
trap 'rm -f "$PROMPT_FILE"' EXIT
|
|
61
|
-
```
|
|
62
|
-
|
|
63
|
-
## Dashboard helper relation
|
|
64
|
-
|
|
65
|
-
When called by `wp dash`, record the outside-voice lifecycle with
|
|
66
|
-
`wp dash-helper-start --provider codex --role reviewer` and
|
|
67
|
-
`wp dash-helper-complete --provider codex --role reviewer --outcome <completed|failed>`
|
|
68
|
-
using the actual terminal outcome and artifact path.
|
|
69
|
-
Standalone calls skip helper emission. Codex native subagent events are not
|
|
70
|
-
parent-linked in the managed lifecycle; this unsupported coverage must stay
|
|
71
|
-
explicit, and completion must never be inferred from text, PIDs, or
|
|
72
|
-
timestamps.
|
|
73
|
-
|
|
74
|
-
## Options
|
|
75
|
-
|
|
76
|
-
Every mode runs through the `wp review run` runtime and honors these environment
|
|
77
|
-
overrides. Set them inline, e.g.
|
|
78
|
-
`CODEX_REVIEW_MODEL=gpt-5.6-sol CODEX_REVIEW_EFFORT=high /codex`.
|
|
79
|
-
|
|
80
|
-
- `CODEX_REVIEW_MODEL` — the Codex model. Defaults to the `model = "..."` in the
|
|
81
|
-
active Codex config (`${CODEX_HOME:-~/.codex}/config.toml`); if neither the
|
|
82
|
-
environment nor config names a model, `wp review run` lets the Codex CLI choose
|
|
83
|
-
its own default. It MUST be a model the current login is provisioned for: ChatGPT-account
|
|
84
|
-
logins use the config's codenamed ids (e.g. `gpt-5.6-sol`) and reject a bare
|
|
85
|
-
`gpt-5.6`. When the user names a model ("review with 5.6-sol"), pass it here.
|
|
86
|
-
- `CODEX_REVIEW_EFFORT` — reasoning effort, default `medium`. The review stage
|
|
87
|
-
accepts only `medium` or `high`; any other value is rejected.
|
|
88
|
-
- `CODEX_REVIEW_IDLE_SECONDS` — idle timeout in seconds for the review stage,
|
|
89
|
-
default `180`.
|
|
90
|
-
- `CODEX_REVIEW_ARTIFACT_ROOT` — directory for private runtime diagnostics,
|
|
91
|
-
default `<cwd>/.webpresso/reviews`. Never commit files from this directory or
|
|
92
|
-
cite them as approval evidence.
|
|
93
|
-
|
|
94
|
-
Direct `wp review run` output is advisory. `wp_review_gate` is plan-only
|
|
95
|
-
exact-version draft approval/auto-promotion; implementation/phase reviews use
|
|
96
|
-
advisory `wp_review_run` and never mutate blueprint approval state. CLI delivery
|
|
97
|
-
gates are MCP-unavailable compatibility only.
|
|
98
|
-
|
|
99
|
-
## Modes
|
|
100
|
-
|
|
101
|
-
### Review
|
|
102
|
-
|
|
103
|
-
1. Capture the current branch, base branch, and `git diff --stat`.
|
|
104
|
-
2. Write a concise prompt asking Codex to find correctness, security, data-loss, and maintainability risks.
|
|
105
|
-
3. Run Codex through the typed review runtime in read-only mode:
|
|
106
|
-
|
|
107
|
-
```bash
|
|
108
|
-
# CODEX_REVIEW_MODEL / CODEX_REVIEW_EFFORT are overridable — see ## Options for
|
|
109
|
-
# the full option list, valid values, and how to pick a login-supported model.
|
|
110
|
-
CODEX_REVIEW_MODEL=${CODEX_REVIEW_MODEL:-}
|
|
111
|
-
CODEX_MODEL_ARGS=()
|
|
112
|
-
if [ -n "$CODEX_REVIEW_MODEL" ]; then
|
|
113
|
-
CODEX_MODEL_ARGS=(--model "$CODEX_REVIEW_MODEL")
|
|
114
|
-
fi
|
|
115
|
-
CODEX_REVIEW_EFFORT=${CODEX_REVIEW_EFFORT:-medium}
|
|
116
|
-
case "$CODEX_REVIEW_EFFORT" in
|
|
117
|
-
medium|high) ;;
|
|
118
|
-
*) echo "CODEX_REVIEW_EFFORT must be one of: medium, high" >&2; exit 2 ;;
|
|
119
|
-
esac
|
|
120
|
-
CODEX_REVIEW_IDLE_SECONDS=${CODEX_REVIEW_IDLE_SECONDS:-180}
|
|
121
|
-
CODEX_REVIEW_ARTIFACT_ROOT=${CODEX_REVIEW_ARTIFACT_ROOT:-"$(pwd)/.webpresso/reviews"}
|
|
122
|
-
CODEX_REVIEW_CODE=0
|
|
123
|
-
wp review run \
|
|
124
|
-
--provider codex \
|
|
125
|
-
--prompt-file "$PROMPT_FILE" \
|
|
126
|
-
"${CODEX_MODEL_ARGS[@]}" \
|
|
127
|
-
--effort "$CODEX_REVIEW_EFFORT" \
|
|
128
|
-
--stage review \
|
|
129
|
-
--artifact-root "$CODEX_REVIEW_ARTIFACT_ROOT" \
|
|
130
|
-
--idle-seconds "$CODEX_REVIEW_IDLE_SECONDS"
|
|
131
|
-
CODEX_REVIEW_CODE=$?
|
|
132
|
-
```
|
|
133
|
-
|
|
134
|
-
4. Summarize findings with severity, evidence, and whether you independently verified them.
|
|
34
|
+
Use authoritative mode only when formal plan or delivery approval is requested.
|
|
135
35
|
|
|
136
|
-
|
|
36
|
+
## Model policy
|
|
137
37
|
|
|
138
|
-
|
|
38
|
+
When the user names an exact model, pass it in the MCP `model` field. Otherwise omit
|
|
39
|
+
that field and let the MCP owner select the configured provider default.
|
|
139
40
|
|
|
140
|
-
|
|
41
|
+
## Review quality
|
|
141
42
|
|
|
142
|
-
|
|
43
|
+
Follow `catalog/agent/rules/review-methodology-sota.md`; report typed verdict, model,
|
|
44
|
+
artifact, grounded findings, and independently verified conclusions.
|
package/skills/grok/SKILL.md
CHANGED
|
@@ -1,32 +1,44 @@
|
|
|
1
1
|
---
|
|
2
2
|
name: grok
|
|
3
|
-
description: "Grok outside-voice
|
|
3
|
+
description: "Grok outside-voice reviewer through Webpresso MCP."
|
|
4
4
|
license: MIT
|
|
5
5
|
---
|
|
6
6
|
|
|
7
|
-
|
|
7
|
+
# Grok outside-voice review
|
|
8
8
|
|
|
9
|
-
-
|
|
10
|
-
|
|
11
|
-
|
|
12
|
-
|
|
13
|
-
|
|
14
|
-
- Call `
|
|
15
|
-
|
|
16
|
-
-
|
|
17
|
-
|
|
9
|
+
Use only for a requested content-bound plan or delivery review. Treat output as external advice
|
|
10
|
+
until independently verified.
|
|
11
|
+
|
|
12
|
+
## MCP-only contract
|
|
13
|
+
|
|
14
|
+
- Call `wp_review_gate` exactly once for the real plan version or delivery diff.
|
|
15
|
+
- Observe the same durable operation with `wp_review_gate_wait` when needed.
|
|
16
|
+
- Never invoke a provider command, the legacy review CLI, or a removed review-run surface.
|
|
17
|
+
- If the review MCP is unavailable, stop and report it unavailable. There is no CLI fallback.
|
|
18
|
+
|
|
19
|
+
## Advisory call
|
|
18
20
|
|
|
19
21
|
```jsonc
|
|
20
22
|
{
|
|
21
|
-
"
|
|
23
|
+
"project_id": "<project>",
|
|
24
|
+
"slug": "<blueprint-slug>",
|
|
25
|
+
"purpose": "delivery",
|
|
26
|
+
"base_ref": "<full-base-sha>",
|
|
27
|
+
"authority_ref": "<full-head-sha>",
|
|
22
28
|
"provider": "grok",
|
|
23
29
|
"repository_access": "none",
|
|
24
|
-
"
|
|
25
|
-
"stage": "review",
|
|
26
|
-
"expectedMarker": "VERDICT:",
|
|
30
|
+
"mode": "advisory",
|
|
27
31
|
}
|
|
28
32
|
```
|
|
29
33
|
|
|
30
|
-
|
|
31
|
-
|
|
32
|
-
|
|
34
|
+
Use authoritative mode only when formal plan or delivery approval is requested.
|
|
35
|
+
|
|
36
|
+
## Model policy
|
|
37
|
+
|
|
38
|
+
When the user names an exact model, pass it in the MCP `model` field. Otherwise omit
|
|
39
|
+
that field and let the MCP owner select the configured provider default.
|
|
40
|
+
|
|
41
|
+
## Review quality
|
|
42
|
+
|
|
43
|
+
Follow `catalog/agent/rules/review-methodology-sota.md`; report typed verdict, model,
|
|
44
|
+
artifact, grounded findings, and independently verified conclusions.
|
|
@@ -55,34 +55,44 @@ Each check prints `[x]` (pass) or `[ ]` (fail) with a detail line:
|
|
|
55
55
|
|
|
56
56
|
## Failure Remediation
|
|
57
57
|
|
|
58
|
-
| Check
|
|
59
|
-
|
|
|
60
|
-
| `pretool-guard` / `post-tool` / etc. — not found
|
|
61
|
-
| `pretool-guard` / etc. — not executable
|
|
62
|
-
| `plugin.json integrity` — missing
|
|
63
|
-
| `MCP server liveness` — timeout
|
|
64
|
-
|
|
|
65
|
-
| Any check — not found at `dist/esm/...` | Build artifacts missing | Run `wp run build` in the webpresso repo |
|
|
58
|
+
| Check | Likely Cause | Fix |
|
|
59
|
+
| ------------------------------------------------ | -------------------------------------------------------------- | ----------------------------------------------------------------- |
|
|
60
|
+
| `pretool-guard` / `post-tool` / etc. — not found | `wp run build` not run after install | `wp run build` |
|
|
61
|
+
| `pretool-guard` / etc. — not executable | `chmod +x` not persisted | Re-run `wp run prepare` or `wp run build` which runs `chmod-bins` |
|
|
62
|
+
| `plugin.json integrity` — missing | Claude adapter manifest absent from `@webpresso/plugin-claude` | Re-run `wp setup --host claude` or reinstall the Claude plugin |
|
|
63
|
+
| `MCP server liveness` — timeout | MCP server cold-start too slow | Wait and retry, or run `wp hooks doctor --skip-mcp` |
|
|
64
|
+
| Any check — not found at `dist/esm/...` | Build artifacts missing | Run `wp run build` in the webpresso repo |
|
|
66
65
|
|
|
67
66
|
### Updating the webpresso MCP server (monorepo authors)
|
|
68
67
|
|
|
69
|
-
`mcp` is **
|
|
70
|
-
|
|
68
|
+
`mcp` is **source-first**: in a checkout with TypeScript source the server runs
|
|
69
|
+
from source, so it always serves HEAD. After changing `src/mcp/**`, conversation
|
|
70
|
+
catalog, or pretool routing that MCP hosts must run:
|
|
71
71
|
|
|
72
|
-
1.
|
|
73
|
-
2.
|
|
74
|
-
`pkill -f 'dist/runtime/.*/wp mcp'` if a long-lived process sticks.
|
|
75
|
-
3. Smoke a tool only available in the new code.
|
|
72
|
+
1. Reconnect webpresso MCP in the agent host.
|
|
73
|
+
2. Smoke a tool only available in the new code.
|
|
76
74
|
|
|
77
|
-
|
|
78
|
-
|
|
79
|
-
|
|
75
|
+
No runtime rebuild is involved, and no `pkill` is needed — every launcher layer
|
|
76
|
+
owns its child's stdin, so a disconnect unwinds the whole chain. Chat-only
|
|
77
|
+
restart is still **not** enough: reconnect the server itself.
|
|
78
|
+
|
|
79
|
+
Managed `wp hook` and `wp hooks doctor|status|dispatch` commands are
|
|
80
|
+
source-authoritative and require no post-merge runtime rebuild. Existing
|
|
81
|
+
managed installations run exact `wp sync`, then rerun `wp hooks doctor` or
|
|
82
|
+
`wp hooks status` until the legacy-projection warning is absent. Packed
|
|
83
|
+
consumers continue to use the shipped compiled runtime and repair/reinstall that
|
|
84
|
+
artifact if it is missing.
|
|
85
|
+
|
|
86
|
+
**One-time upgrade step:** chains started before the source-first lane shipped
|
|
87
|
+
ran no stdin relay, so quit and reopen each MCP host once to retire them. A
|
|
88
|
+
`pkill` pattern is not a substitute — it misses source-launched chains and can
|
|
89
|
+
kill another session's server.
|
|
80
90
|
|
|
81
91
|
**Offline shell allowlist** when webpresso MCP tools are unavailable (must
|
|
82
92
|
match `dev-routing-bounded-read.ts`):
|
|
83
93
|
|
|
84
94
|
```text
|
|
85
|
-
rg -n pattern path | head -
|
|
95
|
+
rg --no-config -n pattern path 2>/dev/null | head -c 65536
|
|
86
96
|
head -n 200 path/to/file
|
|
87
97
|
git log --oneline -n 50
|
|
88
98
|
git show --stat HEAD
|
|
@@ -4,94 +4,44 @@ description: "OpenCode Go aggregate outside-voice reviewer for read-only plan, c
|
|
|
4
4
|
license: MIT
|
|
5
5
|
---
|
|
6
6
|
|
|
7
|
-
# OpenCode Go aggregate
|
|
8
|
-
|
|
9
|
-
Use
|
|
10
|
-
|
|
11
|
-
|
|
12
|
-
|
|
13
|
-
|
|
14
|
-
|
|
15
|
-
|
|
16
|
-
|
|
17
|
-
|
|
18
|
-
|
|
19
|
-
|
|
20
|
-
|
|
21
|
-
|
|
22
|
-
|
|
23
|
-
|
|
24
|
-
|
|
25
|
-
|
|
26
|
-
|
|
27
|
-
|
|
28
|
-
|
|
29
|
-
|
|
30
|
-
|
|
31
|
-
|
|
32
|
-
|
|
33
|
-
trap 'rm -f "$PROMPT_FILE"' EXIT
|
|
34
|
-
```
|
|
35
|
-
|
|
36
|
-
## Dashboard helper relation
|
|
37
|
-
|
|
38
|
-
When `wp dash` provides dashboard context, use the shared helper emitter around this outside-voice review. The commands leave standalone OpenCode Go reviews unchanged.
|
|
39
|
-
|
|
40
|
-
```bash
|
|
41
|
-
DASH_HELPER_RUN_ID=""
|
|
42
|
-
if [ "${WP_DASH:-}" = "1" ] && [ -n "${WP_DASH_RUN_ID:-}" ]; then
|
|
43
|
-
DASH_HELPER_RUN_ID=$(wp dash-helper-start --provider opencode-go --role reviewer --artifact "${OPENCODE_GO_REVIEW_ARTIFACT:-}")
|
|
44
|
-
fi
|
|
7
|
+
# OpenCode Go aggregate outside-voice review
|
|
8
|
+
|
|
9
|
+
Use only for a requested content-bound plan or delivery review. Treat output as external advice
|
|
10
|
+
until independently verified.
|
|
11
|
+
|
|
12
|
+
## MCP-only contract
|
|
13
|
+
|
|
14
|
+
- Call `wp_review_gate` exactly once for the real plan version or delivery diff.
|
|
15
|
+
- Observe the same durable operation with `wp_review_gate_wait` when needed.
|
|
16
|
+
- Never invoke a provider command, the legacy review CLI, or a removed review-run surface.
|
|
17
|
+
- If the review MCP is unavailable, stop and report it unavailable. There is no CLI fallback.
|
|
18
|
+
|
|
19
|
+
## Advisory call
|
|
20
|
+
|
|
21
|
+
```jsonc
|
|
22
|
+
{
|
|
23
|
+
"project_id": "<project>",
|
|
24
|
+
"slug": "<blueprint-slug>",
|
|
25
|
+
"purpose": "delivery",
|
|
26
|
+
"base_ref": "<full-base-sha>",
|
|
27
|
+
"authority_ref": "<full-head-sha>",
|
|
28
|
+
"provider": "opencode",
|
|
29
|
+
"model": "opencode-go/deepseek-v4-pro",
|
|
30
|
+
"repository_access": "none",
|
|
31
|
+
"mode": "advisory",
|
|
32
|
+
}
|
|
45
33
|
```
|
|
46
34
|
|
|
47
|
-
|
|
35
|
+
Use authoritative mode only when formal plan or delivery approval is requested.
|
|
48
36
|
|
|
49
|
-
|
|
50
|
-
OPENCODE_GO_REVIEW_CODE=$?
|
|
51
|
-
if [ -n "$DASH_HELPER_RUN_ID" ]; then
|
|
52
|
-
if [ "$OPENCODE_GO_REVIEW_CODE" -eq 0 ]; then
|
|
53
|
-
wp dash-helper-complete --provider opencode-go --run-id "$DASH_HELPER_RUN_ID" --role reviewer --outcome completed --artifact "${OPENCODE_GO_REVIEW_ARTIFACT:-}"
|
|
54
|
-
else
|
|
55
|
-
wp dash-helper-complete --provider opencode-go --run-id "$DASH_HELPER_RUN_ID" --role reviewer --outcome failed --artifact "${OPENCODE_GO_REVIEW_ARTIFACT:-}"
|
|
56
|
-
fi
|
|
57
|
-
fi
|
|
58
|
-
```
|
|
59
|
-
|
|
60
|
-
The managed OpenCode plugin currently has no parent-linked native subagent completion lifecycle. This capability is explicit unsupported coverage, not an inferred dashboard relation.
|
|
61
|
-
|
|
62
|
-
## MCP-unavailable fallback: review command
|
|
37
|
+
## Model policy
|
|
63
38
|
|
|
64
|
-
|
|
65
|
-
|
|
39
|
+
Default exact model: `opencode-go/deepseek-v4-pro`.
|
|
40
|
+
Committed exact options: `opencode-go/deepseek-v4-pro`, `opencode-go/deepseek-v4-flash`, `opencode-go/kimi-k2.7-code`, `opencode-go/glm-5.3`, `opencode-go/minimax-m3`, `opencode-go/mimo-v2.5-pro`, `opencode-go/qwen3.8-max`, `opencode-go/mimo-v2.5`, `opencode-go/qwen3.7-plus`, `opencode-go/qwen3.7-max`, `opencode-go/qwen3.6-plus`, `opencode-go/gpt-5.6-luna`, `opencode-go/minimax-m2.7`, `opencode-go/kimi-k3`, `opencode-go/glm-5.2`, `opencode-go/kimi-k2.6`, `opencode-go/hy3`, `opencode-go/glm-5.1`, `opencode-go/grok-4.5`.
|
|
41
|
+
Use a named exact option when requested; otherwise use the default. Provider catalog discovery
|
|
42
|
+
belongs to the snapshot drift workflow, never to a skill invocation.
|
|
66
43
|
|
|
67
|
-
|
|
68
|
-
# Aggregate reviewer: honor the canonical model-selection policy (Kimi K2.7 Code → DeepSeek V4 Pro → DeepSeek V4 Flash first, then cross-family selection fallbacks).
|
|
69
|
-
CATALOG=$(opencode models opencode-go)
|
|
70
|
-
MODEL=$(echo "$CATALOG" | grep '^opencode-go/kimi' | grep -- '-code$' | sort -V | tail -1)
|
|
71
|
-
[ -z "$MODEL" ] && MODEL=$(echo "$CATALOG" | grep '^opencode-go/deepseek' | grep -- '-pro$' | sort -V | tail -1)
|
|
72
|
-
[ -z "$MODEL" ] && MODEL=$(echo "$CATALOG" | grep '^opencode-go/deepseek' | grep -- '-flash$' | sort -V | tail -1)
|
|
73
|
-
[ -z "$MODEL" ] && MODEL=$(echo "$CATALOG" | grep '^opencode-go/qwen' | grep -- '-max$' | sort -V | tail -1)
|
|
74
|
-
[ -z "$MODEL" ] && MODEL=$(echo "$CATALOG" | grep '^opencode-go/glm' | sort -V | tail -1)
|
|
75
|
-
[ -z "$MODEL" ] && MODEL=$(echo "$CATALOG" | grep '^opencode-go/minimax' | sort -V | tail -1)
|
|
76
|
-
[ -z "$MODEL" ] && MODEL=$(echo "$CATALOG" | grep '^opencode-go/mimo' | grep -- '-pro$' | sort -V | tail -1)
|
|
77
|
-
[ -z "$MODEL" ] && MODEL=$(echo "$CATALOG" | sort -V | tail -1)
|
|
78
|
-
|
|
79
|
-
# Resolve one model from one bounded catalog read, then launch one review.
|
|
80
|
-
[ -n "$MODEL" ] || { echo "No OpenCode Go model resolved for this reviewer." >&2; exit 2; }
|
|
81
|
-
OPENCODE_GO_REVIEW_EFFORT=${OPENCODE_GO_REVIEW_EFFORT:-medium}
|
|
82
|
-
[ "$OPENCODE_GO_REVIEW_EFFORT" = medium ] || [ "$OPENCODE_GO_REVIEW_EFFORT" = high ] || { echo "OPENCODE_GO_REVIEW_EFFORT must be one of: medium, high" >&2; exit 2; }
|
|
83
|
-
OPENCODE_GO_REVIEW_IDLE_SECONDS=${OPENCODE_GO_REVIEW_IDLE_SECONDS:-180}
|
|
84
|
-
OPENCODE_GO_REVIEW_ARTIFACT_ROOT=${OPENCODE_GO_REVIEW_ARTIFACT_ROOT:-"$(pwd)/.webpresso/reviews"}
|
|
85
|
-
wp review run \
|
|
86
|
-
--provider opencode \
|
|
87
|
-
--prompt-file "$PROMPT_FILE" \
|
|
88
|
-
--model "$MODEL" \
|
|
89
|
-
--effort "$OPENCODE_GO_REVIEW_EFFORT" \
|
|
90
|
-
--stage review \
|
|
91
|
-
--artifact-root "$OPENCODE_GO_REVIEW_ARTIFACT_ROOT" \
|
|
92
|
-
--idle-seconds "$OPENCODE_GO_REVIEW_IDLE_SECONDS"
|
|
93
|
-
```
|
|
44
|
+
## Review quality
|
|
94
45
|
|
|
95
|
-
|
|
96
|
-
|
|
97
|
-
OpenCode Go usage limit ends all OpenCode models this review (no family hop). Next: Grok (`--provider grok`). Gate fails over to Grok when accounts cool. Treat `true-idle`, `protocol-unsupported`, provider/abort/spawn/artifact failure as unavailable. Never replace with static timeout, byte-growth heartbeat, or buffered-output fallback.
|
|
46
|
+
Follow `catalog/agent/rules/review-methodology-sota.md`; report typed verdict, model,
|
|
47
|
+
artifact, grounded findings, and independently verified conclusions.
|
|
@@ -77,6 +77,7 @@ Run **1–3 (and relevant review lenses) in parallel** whenever tools allow. Do
|
|
|
77
77
|
4. **Plan-review lenses (skip when the caller already ran them).** Folded `/autoplan`: run relevant CEO/founder, design, engineering, and DevEx lenses **in parallel when independent**; consolidate keep/change/drop, unresolved taste calls, tests, go/no-go.
|
|
78
78
|
- **Skip condition.** If the caller states `plan-* reviews complete for this pass`, do **not** re-run the lenses. `ralplan` runs them before handing off, so re-running them here is duplicated spend, not a second opinion.
|
|
79
79
|
- **Re-entry.** If refinement itself **materially changes scope**, re-run only the affected lens — never the full set — then continue applying.
|
|
80
|
+
- **Re-review rounds.** `catalog/agent/rules/review-methodology-sota.md` Principle 5 sets the delta-anchored adjudication contract subsequent `wp_review_gate` rounds follow.
|
|
80
81
|
|
|
81
82
|
5. **Blueprint enforcement (max parallel).**
|
|
82
83
|
- Split for independent execution; declare `Depends`; file-conflict CP = 0 per wave; TDD/proof steps; preserve acceptance criteria.
|
|
@@ -333,18 +333,18 @@ without reading other tasks. Include: what, why, constraints, gotchas.]
|
|
|
333
333
|
**Steps (TDD):**
|
|
334
334
|
|
|
335
335
|
1. Write failing test for [specific behavior]
|
|
336
|
-
2. Run: `
|
|
336
|
+
2. Run: `wp test --files <path/to/test-file.test.ts>` — verify the current contract
|
|
337
337
|
3. Implement minimal code to pass
|
|
338
|
-
4. Run: `
|
|
338
|
+
4. Run: `wp test --files <path/to/test-file.test.ts>` — verify PASS
|
|
339
339
|
5. Refactor if needed (complexity ≤ 8)
|
|
340
|
-
6. Run: `
|
|
340
|
+
6. Run: `wp lint --files <changed-file.ts> <changed-test.ts>` and `wp typecheck --files <changed-file.ts> <changed-test.ts>`
|
|
341
341
|
|
|
342
342
|
**Acceptance:**
|
|
343
343
|
|
|
344
|
-
- [ ]
|
|
344
|
+
- [ ] Positive contract test covers the requested behavior
|
|
345
345
|
- [ ] Implementation passes all tests
|
|
346
|
-
- [ ] `
|
|
347
|
-
- [ ] `
|
|
346
|
+
- [ ] `wp lint --files <changed-files...>` passes
|
|
347
|
+
- [ ] `wp typecheck --files <changed-files...>` passes
|
|
348
348
|
```
|
|
349
349
|
|
|
350
350
|
Use `#### Task X.Y: ...` only when a lane prefix would add no value, but prefer lane-prefixed headers such as `[schema]`, `[backend]`, `[ui]`, `[infra]`, `[docs]`, or `[qa]`.
|
|
@@ -362,22 +362,22 @@ These are enforced project conventions (with webpresso's conventions as the exam
|
|
|
362
362
|
|
|
363
363
|
- Task says "1 day", "3 hours", "2 weeks" → Change to t-shirt size
|
|
364
364
|
- Task creates migration files or migration infrastructure when the repo prefers `db push` → Use the repo's chosen workflow
|
|
365
|
-
- References
|
|
365
|
+
- References a retired database command when the repo uses `db push` → Use `db push` (entity YAML → schema generation → push)
|
|
366
366
|
|
|
367
367
|
### Blueprint Validation Checklist
|
|
368
368
|
|
|
369
369
|
Run this audit on every task in the blueprint:
|
|
370
370
|
|
|
371
|
-
| Check | Violation | Fix
|
|
372
|
-
| --------------------------------- | ---------------------------------------------------- |
|
|
373
|
-
| Has `**Depends:**` line? | Missing → parallel execution can't build DAG | Add explicit dependency or "None"
|
|
374
|
-
| Has `**Files:**` section? | Missing → agents can't detect file conflicts | List every file touched (Create/Modify)
|
|
375
|
-
| Has `**Steps (TDD):**`? | Missing → agents skip tests | Add
|
|
376
|
-
| Has `**Acceptance:**` checkboxes? | Missing → no completion criteria | Add testable acceptance criteria
|
|
377
|
-
| Description self-contained? | References "see above" or "as described in Task X.Y" | Inline the context — each task runs independently
|
|
378
|
-
| Files overlap with another task? | Two tasks modify same file → conflict in parallel | Merge tasks or add explicit `**Depends:**`
|
|
379
|
-
| Uses t-shirt sizing? | Day/week estimates used | Replace with XS/S/M/L/XL
|
|
380
|
-
| Follows repo DB workflow? | Diverges from repo's chosen workflow | Use repo's chosen workflow instead
|
|
371
|
+
| Check | Violation | Fix |
|
|
372
|
+
| --------------------------------- | ---------------------------------------------------- | -------------------------------------------------------- |
|
|
373
|
+
| Has `**Depends:**` line? | Missing → parallel execution can't build DAG | Add explicit dependency or "None" |
|
|
374
|
+
| Has `**Files:**` section? | Missing → agents can't detect file conflicts | List every file touched (Create/Modify) |
|
|
375
|
+
| Has `**Steps (TDD):**`? | Missing → agents skip tests | Add positive verification steps with exact `wp` commands |
|
|
376
|
+
| Has `**Acceptance:**` checkboxes? | Missing → no completion criteria | Add testable acceptance criteria |
|
|
377
|
+
| Description self-contained? | References "see above" or "as described in Task X.Y" | Inline the context — each task runs independently |
|
|
378
|
+
| Files overlap with another task? | Two tasks modify same file → conflict in parallel | Merge tasks or add explicit `**Depends:**` |
|
|
379
|
+
| Uses t-shirt sizing? | Day/week estimates used | Replace with XS/S/M/L/XL |
|
|
380
|
+
| Follows repo DB workflow? | Diverges from repo's chosen workflow | Use repo's chosen workflow instead |
|
|
381
381
|
|
|
382
382
|
### Granularity Rules
|
|
383
383
|
|
|
@@ -532,7 +532,7 @@ Rate the plan's parallelizability:
|
|
|
532
532
|
|
|
533
533
|
### Self-Contained Task Test
|
|
534
534
|
|
|
535
|
-
For each task, ask: **"Can an agent execute this task with ONLY the task description, the codebase, and `
|
|
535
|
+
For each task, ask: **"Can an agent execute this task with ONLY the task description, the codebase, and the repository's `wp` commands?"**
|
|
536
536
|
|
|
537
537
|
If the answer is no, the task is missing context. Common fixes:
|
|
538
538
|
|
package/skills/ralplan/SKILL.md
CHANGED
|
@@ -112,6 +112,7 @@ Rules:
|
|
|
112
112
|
- Apply findings **into the blueprint** after each review (or once after the set if independent).
|
|
113
113
|
- Independent reviews may run in parallel **only when** they do not depend on each other's edits; if eng findings reshape scope, re-run only the affected optional reviews.
|
|
114
114
|
- Optional high-risk challenge: one sequential outside-voice pass (`codex` / `claude` / `opencode-go`) **after** eng review if the user asks or the change is auth / data-loss / public-API sensitive. Never parallel multi-host review.
|
|
115
|
+
- Any formal `wp_review_gate` re-review round follows `catalog/agent/rules/review-methodology-sota.md` Principle 5's delta-anchored adjudication contract, not a fresh rediscovery pass.
|
|
115
116
|
|
|
116
117
|
### 5. `/plan-refine` last (always)
|
|
117
118
|
|
|
@@ -34,6 +34,12 @@ wp tech-debt new "Legacy CLI complexity" \
|
|
|
34
34
|
# Preview without writing
|
|
35
35
|
wp tech-debt new "Performance bottleneck" --severity high --category mutation --dry-run
|
|
36
36
|
|
|
37
|
+
# `new` refuses on a primary-like checkout: a file written there cannot be
|
|
38
|
+
# edited afterwards under worktree discipline, so the capture is left half-done.
|
|
39
|
+
# Retry from a managed worktree, point --cwd at one, or override deliberately.
|
|
40
|
+
wp tech-debt new "Legacy CLI complexity" --cwd /path/to/managed/worktree
|
|
41
|
+
wp tech-debt new "Legacy CLI complexity" --allow-primary
|
|
42
|
+
|
|
37
43
|
# List all tech-debt items (optional filters)
|
|
38
44
|
wp tech-debt list
|
|
39
45
|
wp tech-debt list --status accepted
|
|
@@ -434,7 +434,7 @@ Before claiming a test is "done":
|
|
|
434
434
|
- [ ] **Does it use real dependencies?** (PGlite for DB, real services)
|
|
435
435
|
- [ ] **Are assertions specific?** (Not just `toBeTruthy()`)
|
|
436
436
|
- [ ] **Does it test behavior, not implementation?** (No spy assertions)
|
|
437
|
-
- [ ] **Mutation score ≥85%?** (Run `
|
|
437
|
+
- [ ] **Mutation score ≥85%?** (Run `wp test --mutation --package <pkg>`)
|
|
438
438
|
- [ ] **Does it fail if the code breaks?** (Temporarily break code, verify test fails)
|
|
439
439
|
- [ ] **Is it in the right file?** (`.test.ts` for unit, `.integration.test.ts` for DB)
|
|
440
440
|
|
|
@@ -450,23 +450,19 @@ Before claiming a test is "done":
|
|
|
450
450
|
|
|
451
451
|
## Quick Commands
|
|
452
452
|
|
|
453
|
-
These
|
|
453
|
+
These use the repository's `wp` task facade.
|
|
454
454
|
|
|
455
455
|
```bash
|
|
456
456
|
# Run tests
|
|
457
457
|
# WARNING: Never run full suites during iteration. Use single-file verification.
|
|
458
|
-
|
|
459
|
-
|
|
460
|
-
|
|
461
|
-
|
|
462
|
-
# Mutation testing
|
|
463
|
-
just test --mutation --package <package> # Full mutation test
|
|
464
|
-
just test --mutation-diff # Changed packages only
|
|
458
|
+
wp test # All tests (FINAL VERIFICATION ONLY)
|
|
459
|
+
wp test --package <package> # Specific package (FINAL VERIFICATION ONLY)
|
|
460
|
+
wp test --files path/to/test.ts # Single file (ITERATION SAFE)
|
|
465
461
|
|
|
466
462
|
# Audit quality
|
|
467
|
-
|
|
468
|
-
|
|
469
|
-
|
|
463
|
+
wp audit test-smells # Static test-quality checks
|
|
464
|
+
wp audit tph-e2e # Integration-first E2E review
|
|
465
|
+
wp qa # Full quality check
|
|
470
466
|
```
|
|
471
467
|
|
|
472
468
|
## Decision Tree
|