continuous-improvement 3.15.0 → 3.16.0
This diff represents the content of publicly available package versions that have been released to one of the supported registries. The information contained in this diff is provided for informational purposes only and reflects changes between package versions as they appear in their respective public registries.
- package/.claude-plugin/marketplace.json +2 -2
- package/CHANGELOG.md +14 -0
- package/README.md +4 -3
- package/bin/check-skill-count-prose.mjs +168 -0
- package/bin/install.mjs +46 -1
- package/bin/lint-transcript.mjs +15 -3
- package/bin/mcp-server.mjs +1 -0
- package/commands/roast.md +34 -0
- package/hooks/workflow-distill.mjs +145 -0
- package/lib/plugin-metadata.mjs +8 -3
- package/lib/version-check.mjs +115 -0
- package/llms.txt +1 -1
- package/package.json +4 -3
- package/plugins/beginner.json +1 -1
- package/plugins/continuous-improvement/.claude-plugin/marketplace.json +2 -2
- package/plugins/continuous-improvement/.claude-plugin/plugin.json +2 -2
- package/plugins/continuous-improvement/bin/mcp-server.mjs +1 -0
- package/plugins/continuous-improvement/commands/roast.md +34 -0
- package/plugins/continuous-improvement/hooks/hooks.json +6 -1
- package/plugins/continuous-improvement/hooks/workflow-distill.mjs +145 -0
- package/plugins/continuous-improvement/lib/plugin-metadata.mjs +8 -3
- package/plugins/continuous-improvement/skills/README.md +1 -0
- package/plugins/continuous-improvement/skills/roast/SKILL.md +108 -0
- package/plugins/continuous-improvement/skills/strategic-compact/SKILL.md +12 -32
- package/plugins/expert.json +1 -1
- package/skills/README.md +1 -1
- package/skills/roast.md +108 -0
- package/skills/strategic-compact.md +12 -32
package/skills/README.md
CHANGED
|
@@ -33,7 +33,7 @@ Tier-2 skills layer on top of tier-1 for users running `npx continuous-improveme
|
|
|
33
33
|
|-------|--------------|------------------|
|
|
34
34
|
| `safety-guard` | Three-mode runtime guard (careful/freeze/guard) that blocks destructive commands and locks edits to a directory | Autonomous loops, prod systems, `--dangerously-skip-permissions` sessions |
|
|
35
35
|
| `token-budget-advisor` | Heuristic input/output token estimator that offers 25%/50%/75%/100% depth choices before answering | Long sessions where response size matters |
|
|
36
|
-
| `strategic-compact` |
|
|
36
|
+
| `strategic-compact` | Manual phase-boundary checklist for deciding when to run `/compact` (research→plan, plan→implement, debug→next) instead of relying on arbitrary auto-compaction | Multi-phase tasks that approach context limits |
|
|
37
37
|
| `wild-risa-balance` | Decision-framing lens that pairs WILD (Wild/Imaginative/Limitless/Disruptive) generation with RISA (Realistic/Important/Specific/Agreeable) execution, used to split recommendation lists into bold pilots above a safe baseline | Multi-item recommendation blocks where bold options keep losing to safe ones in a flat list |
|
|
38
38
|
|
|
39
39
|
The `/learn-eval` slash command also ships as part of the expert install: extract a session pattern, run a checklist quality gate, and decide global-vs-project save location before writing any skill file.
|
package/skills/roast.md
ADDED
|
@@ -0,0 +1,108 @@
|
|
|
1
|
+
---
|
|
2
|
+
name: roast
|
|
3
|
+
tier: "2"
|
|
4
|
+
description: Enforces Law 1 (Research Before Executing) of the 7 Laws of AI Agent Discipline. Convene a 5-persona adversarial council (Contrarian, Expansionist, Logician, Researcher, Buyer) that attacks an idea from every angle, then a Judge returns one GO / RESHAPE / KILL verdict plus the cheapest 48-hour test to de-risk it — so you pressure-test an idea before sinking time into building the wrong thing.
|
|
5
|
+
origin: continuous-improvement
|
|
6
|
+
user-invocable: true
|
|
7
|
+
argument-hint: "[the idea to roast]"
|
|
8
|
+
---
|
|
9
|
+
|
|
10
|
+
# /roast — Convene the council before you build
|
|
11
|
+
|
|
12
|
+
Claude's default is to agree with you. `/roast` is the opposite. Law 1 says research before executing — and the cheapest, most-skipped research is an honest adversarial read of the idea *itself* before any plan or code exists. This skill convenes a council of five independent persona agents who tear an idea apart and build it up from every angle, then a Judge synthesizes everything into one decisive verdict.
|
|
13
|
+
|
|
14
|
+
The council is adversarial on purpose. No persona is allowed to hedge or be polite. The point is to surface what you can't see because you're too close to it — and to do it in an hour, not after a month of building.
|
|
15
|
+
|
|
16
|
+
Adapted from the community `/roast` idea-council pattern and reshaped as a continuous-improvement-native Law 1 skill.
|
|
17
|
+
|
|
18
|
+
## When to activate
|
|
19
|
+
|
|
20
|
+
- Before you sink time or money into building something — a product, a feature, a business, a bet.
|
|
21
|
+
- When you catch yourself (or the agent) agreeing with a plan that has never been attacked.
|
|
22
|
+
- The user types `/roast`, "roast this idea", "convene the council", "pressure-test this", "stress-test this idea", "validate this business idea", or "give me a brutal second opinion".
|
|
23
|
+
- A `/proceed-with-the-recommendation` walk is about to start but the *premise* underneath the recommendation list was never challenged.
|
|
24
|
+
|
|
25
|
+
## Step 1: Get the brief
|
|
26
|
+
|
|
27
|
+
If `$ARGUMENTS` contains the idea, start there. Then ask a tight set of clarifying questions so the council judges something real. Ask only what hasn't already been provided — 3-4 questions max, in **one batch**:
|
|
28
|
+
|
|
29
|
+
1. **The idea** in one or two sentences (what it is, what it does).
|
|
30
|
+
2. **Who it's for** and **how it makes money** (the buyer + the price/model).
|
|
31
|
+
3. **Your edge** — relevant skills, audience, or assets you already have.
|
|
32
|
+
4. **Constraints** — budget, timeline, how fast you need the first dollar.
|
|
33
|
+
|
|
34
|
+
If the user says "just run it" or has already given you enough, skip the questions and proceed. Don't over-interrogate — one round, then convene.
|
|
35
|
+
|
|
36
|
+
Write the brief into a single short paragraph you will paste verbatim into every council member's prompt, so all five judge the same thing.
|
|
37
|
+
|
|
38
|
+
## Step 2: Convene the council (5 agents, in parallel)
|
|
39
|
+
|
|
40
|
+
Spin up **all five agents in parallel in a single message** — one subagent each (`general-purpose`). Paste the same brief into each, then give it its persona mandate below.
|
|
41
|
+
|
|
42
|
+
Each council member must return: a one-line stance, their 3-5 sharpest points, the single most important thing the user must hear, and a 1-10 score on their own dimension (1 = walk away, 10 = no-brainer).
|
|
43
|
+
|
|
44
|
+
**1. The Contrarian (Red Team)**
|
|
45
|
+
> You are the Contrarian on an idea council. Assume this idea fails. Find the fatal flaws, the fastest way it dies, and the load-bearing assumptions that are probably wrong. Be ruthless and specific. No hedging, no "but it could work." Attack the weakest points. THE BRIEF: [brief]
|
|
46
|
+
|
|
47
|
+
**2. The Expansionist (Bull)**
|
|
48
|
+
> You are the Expansionist on an idea council. Make the strongest possible case FOR this idea. Find the biggest upside, the 10x version, the adjacent opportunities and unlock points the founder isn't seeing. Fight for the potential. Be specific about where the real money and leverage could be. THE BRIEF: [brief]
|
|
49
|
+
|
|
50
|
+
**3. The Logician (First principles)**
|
|
51
|
+
> You are the Logician on an idea council. Use NO outside research and NO web. Reason purely from first principles: does the core mechanism make sense, do the incentives line up, is the underlying logic sound, does the math even work in theory? Strip it to fundamentals and tell us if it holds together. THE BRIEF: [brief]
|
|
52
|
+
|
|
53
|
+
**4. The Researcher (Evidence)**
|
|
54
|
+
> You are the Researcher on an idea council. Use web search. Bring real-world evidence: who the existing competitors are, market size or demand signals, what comparable products charge, whether this is validated by what's already out there or contradicted by it. Cite what you find. Is the real world saying yes or no? THE BRIEF: [brief]
|
|
55
|
+
|
|
56
|
+
**5. The Buyer (Voice of customer)**
|
|
57
|
+
> You are the Buyer on an idea council. Role-play the exact target customer described in the brief. React as them, in first person. Would you actually pay for this? What's your real objection? What would make you choose a competitor or just do nothing instead? What price feels right, and what would make you say yes today? Be the honest, slightly skeptical customer, not a cheerleader. THE BRIEF: [brief]
|
|
58
|
+
|
|
59
|
+
## Step 3: The Judge delivers the verdict
|
|
60
|
+
|
|
61
|
+
Once all five return, YOU act as the Judge. Read every council member's findings, weigh them, and synthesize one decisive verdict. Do not average the scores. Name the real tension between the personas and resolve it.
|
|
62
|
+
|
|
63
|
+
Fold in the **economics lens** yourself: rough pricing, realistic time-to-first-dollar, and whether the user can actually ship this fast given the edge they described.
|
|
64
|
+
|
|
65
|
+
Output the verdict in this exact shape:
|
|
66
|
+
|
|
67
|
+
```
|
|
68
|
+
## THE VERDICT: GO / RESHAPE / KILL
|
|
69
|
+
Confidence: [low / medium / high]
|
|
70
|
+
|
|
71
|
+
**The call in one line:** [the decision, plainly]
|
|
72
|
+
|
|
73
|
+
**Why:** [2-3 sentences resolving the council's tension]
|
|
74
|
+
|
|
75
|
+
**Biggest risk:** [the single thing most likely to kill it]
|
|
76
|
+
**Biggest upside:** [the strongest reason to do it]
|
|
77
|
+
|
|
78
|
+
**Money read:** [rough price, time-to-first-dollar, can they ship fast]
|
|
79
|
+
|
|
80
|
+
**The cheapest 48-hour test:** [the smallest, fastest thing they can do
|
|
81
|
+
to validate the riskiest assumption BEFORE building anything]
|
|
82
|
+
|
|
83
|
+
**If RESHAPE:** [the specific pivot that fixes the fatal flaw while keeping the upside]
|
|
84
|
+
```
|
|
85
|
+
|
|
86
|
+
Then list the five council scores in one line: `Contrarian X/10 · Expansionist X/10 · Logician X/10 · Researcher X/10 · Buyer X/10`.
|
|
87
|
+
|
|
88
|
+
## Rules
|
|
89
|
+
|
|
90
|
+
- Every persona stays in character. None of them hedges or softens. The value is in the friction.
|
|
91
|
+
- The Judge must make an actual call. "It depends" is not a verdict. Pick GO, RESHAPE, or KILL and own it.
|
|
92
|
+
- The cheapest 48-hour test is the most important output. It's how the user finds out if they're right without building the whole thing.
|
|
93
|
+
- Keep the final verdict skimmable. The council does the depth; the Judge does the decision.
|
|
94
|
+
|
|
95
|
+
## How it fits the 7 Laws
|
|
96
|
+
|
|
97
|
+
| Law | Role of this skill |
|
|
98
|
+
|---|---|
|
|
99
|
+
| Law 1 (Research Before Executing) | The council **is** the research — five independent investigations of an idea's viability before a single line of code is written. |
|
|
100
|
+
| Law 2 (Plan Is Sacred) | The verdict's RESHAPE pivot and cheapest-test become the plan's first checkpoint instead of an invented default. |
|
|
101
|
+
| Law 4 (Verify Before Reporting) | The Judge must commit to a falsifiable GO / RESHAPE / KILL call — the anti-pattern is the hedge, "it depends." |
|
|
102
|
+
|
|
103
|
+
## Pairs with
|
|
104
|
+
|
|
105
|
+
- [`grill-me`](./grill-me.md) — once roast says GO or RESHAPE, `grill-me` hardens the **plan**; roast validates the **idea**. Roast first, then grill.
|
|
106
|
+
- [`proceed-with-the-recommendation`](./proceed-with-the-recommendation.md) — walk the verdict's next steps (the cheapest test, the RESHAPE pivot) top-to-bottom under the 7 Laws.
|
|
107
|
+
- [`wild-risa-balance`](./wild-risa-balance.md) — the verdict is a recommendation; run it through the R-I-S-A filter before acting.
|
|
108
|
+
- [`gateguard`](./gateguard.md) — the runtime gate (`hooks/gateguard.mjs`) that fires when the validated idea finally turns into Edit/Write/Bash.
|
|
@@ -32,37 +32,17 @@ Strategic compaction at logical boundaries:
|
|
|
32
32
|
|
|
33
33
|
## How It Works
|
|
34
34
|
|
|
35
|
-
|
|
36
|
-
|
|
37
|
-
1. **
|
|
38
|
-
2. **
|
|
39
|
-
3. **
|
|
40
|
-
|
|
41
|
-
|
|
42
|
-
|
|
43
|
-
|
|
44
|
-
|
|
45
|
-
|
|
46
|
-
{
|
|
47
|
-
"hooks": {
|
|
48
|
-
"PreToolUse": [
|
|
49
|
-
{
|
|
50
|
-
"matcher": "Edit",
|
|
51
|
-
"hooks": [{ "type": "command", "command": "node ~/.claude/skills/strategic-compact/suggest-compact.js" }]
|
|
52
|
-
},
|
|
53
|
-
{
|
|
54
|
-
"matcher": "Write",
|
|
55
|
-
"hooks": [{ "type": "command", "command": "node ~/.claude/skills/strategic-compact/suggest-compact.js" }]
|
|
56
|
-
}
|
|
57
|
-
]
|
|
58
|
-
}
|
|
59
|
-
}
|
|
60
|
-
```
|
|
61
|
-
|
|
62
|
-
## Configuration
|
|
63
|
-
|
|
64
|
-
Environment variables:
|
|
65
|
-
- `COMPACT_THRESHOLD` — Tool calls before first suggestion (default: 50)
|
|
35
|
+
This skill is a manual phase-boundary checklist, not a bundled runtime hook. Use it when planning or reviewing a long session:
|
|
36
|
+
|
|
37
|
+
1. **Name the current phase** — research, planning, implementation, testing, debugging, release, or handoff.
|
|
38
|
+
2. **Check the next transition** — decide whether the next phase needs fresh context or the current context is still load-bearing.
|
|
39
|
+
3. **Preserve state first** — write the plan, todo list, findings, or handoff note that must survive compaction.
|
|
40
|
+
4. **Compact only at a boundary** — if compaction helps, run `/compact` with a specific summary for the next phase.
|
|
41
|
+
5. **Resume from durable artifacts** — after compaction, re-read the plan/files instead of relying on lost conversation context.
|
|
42
|
+
|
|
43
|
+
## Runtime Boundary
|
|
44
|
+
|
|
45
|
+
The current plugin does not ship `strategic-compact` PreToolUse automation or a threshold script. Treat compaction as an operator/agent decision: this skill gives the decision guide, while Claude Code's native `/compact` command performs the actual compaction.
|
|
66
46
|
|
|
67
47
|
## Compaction Decision Guide
|
|
68
48
|
|
|
@@ -94,7 +74,7 @@ Understanding what persists helps you compact with confidence:
|
|
|
94
74
|
1. **Compact after planning** — Once plan is finalized in TodoWrite, compact to start fresh
|
|
95
75
|
2. **Compact after debugging** — Clear error-resolution context before continuing
|
|
96
76
|
3. **Don't compact mid-implementation** — Preserve context for related changes
|
|
97
|
-
4. **
|
|
77
|
+
4. **Use the checklist** — The phase table helps decide *when*; you still decide *if*
|
|
98
78
|
5. **Write before compacting** — Save important context to files or memory before compacting
|
|
99
79
|
6. **Use `/compact` with a summary** — Add a custom message: `/compact Focus on implementing auth middleware next`
|
|
100
80
|
|