liteagents 2.8.3 → 2.10.0

This diff represents the content of publicly available package versions that have been released to one of the supported registries. The information contained in this diff is provided for informational purposes only and reflects changes between package versions as they appear in their respective public registries.
Files changed (113) hide show
  1. package/CHANGELOG.md +51 -0
  2. package/README.md +25 -31
  3. package/installer/cli.js +6 -2
  4. package/package.json +3 -3
  5. package/packages/ampcode/AGENT.md +9 -15
  6. package/packages/ampcode/agents/code-developer.md +11 -12
  7. package/packages/ampcode/agents/quality-assurance.md +1 -1
  8. package/packages/{droid/commands/systematic-debugging.md → ampcode/commands/debug-method.md} +9 -9
  9. package/packages/ampcode/commands/diff-review.md +78 -0
  10. package/packages/ampcode/commands/friction/friction.js +348 -122
  11. package/packages/ampcode/commands/optimize.md +45 -4
  12. package/packages/ampcode/commands/refactor.md +33 -1
  13. package/packages/ampcode/commands/remember.md +107 -46
  14. package/packages/ampcode/commands/security.md +28 -1
  15. package/packages/ampcode/commands/stash.md +7 -0
  16. package/packages/{droid/commands/test-driven-development.md → ampcode/commands/tdd-flow.md} +2 -2
  17. package/packages/ampcode/commands/test-generate.md +64 -15
  18. package/packages/{opencode/command/testing-anti-patterns.md → ampcode/commands/test-traps.md} +77 -3
  19. package/packages/{droid/commands/root-cause-tracing.md → ampcode/commands/trace-back.md} +3 -3
  20. package/packages/ampcode/commands/{verification-before-completion.md → verify-done.md} +3 -3
  21. package/packages/claude/CLAUDE.md +9 -15
  22. package/packages/claude/agents/code-developer.md +11 -12
  23. package/packages/claude/agents/quality-assurance.md +1 -1
  24. package/packages/claude/commands/diff-review.md +78 -0
  25. package/packages/claude/commands/friction/friction.js +348 -122
  26. package/packages/claude/commands/optimize.md +45 -4
  27. package/packages/claude/commands/refactor.md +33 -1
  28. package/packages/claude/commands/remember.md +107 -46
  29. package/packages/claude/commands/security.md +28 -1
  30. package/packages/claude/commands/stash.md +7 -0
  31. package/packages/claude/commands/test-generate.md +64 -15
  32. package/packages/claude/plugins/live-canvas-marketplace/plugins/live-canvas-channel/package-lock.json +3 -3
  33. package/packages/claude/skills/{systematic-debugging → debug-method}/CREATION-LOG.md +1 -1
  34. package/packages/claude/skills/{systematic-debugging → debug-method}/SKILL.md +9 -9
  35. package/packages/claude/skills/{systematic-debugging → debug-method}/test-academic.md +1 -1
  36. package/packages/claude/skills/{systematic-debugging → debug-method}/test-pressure-1.md +1 -1
  37. package/packages/claude/skills/{systematic-debugging → debug-method}/test-pressure-2.md +1 -1
  38. package/packages/claude/skills/{systematic-debugging → debug-method}/test-pressure-3.md +1 -1
  39. package/packages/claude/skills/{test-driven-development → tdd-flow}/SKILL.md +3 -3
  40. package/packages/claude/skills/{testing-anti-patterns → test-traps}/SKILL.md +77 -3
  41. package/packages/claude/skills/{root-cause-tracing → trace-back}/SKILL.md +3 -3
  42. package/packages/claude/skills/{verification-before-completion → verify-done}/SKILL.md +3 -3
  43. package/packages/droid/AGENTS.md +8 -14
  44. package/packages/{opencode/command/systematic-debugging.md → droid/commands/debug-method.md} +9 -9
  45. package/packages/droid/commands/diff-review.md +78 -0
  46. package/packages/droid/commands/friction/friction.js +348 -122
  47. package/packages/droid/commands/optimize.md +45 -4
  48. package/packages/droid/commands/refactor.md +33 -1
  49. package/packages/droid/commands/remember.md +107 -46
  50. package/packages/droid/commands/security.md +28 -1
  51. package/packages/droid/commands/stash.md +7 -0
  52. package/packages/{opencode/command/test-driven-development.md → droid/commands/tdd-flow.md} +2 -2
  53. package/packages/droid/commands/test-generate.md +64 -15
  54. package/packages/droid/commands/{testing-anti-patterns.md → test-traps.md} +77 -3
  55. package/packages/{opencode/command/root-cause-tracing.md → droid/commands/trace-back.md} +3 -3
  56. package/packages/droid/commands/{verification-before-completion.md → verify-done.md} +3 -3
  57. package/packages/droid/droids/code-developer.md +11 -12
  58. package/packages/droid/droids/quality-assurance.md +1 -1
  59. package/packages/opencode/AGENTS.md +8 -14
  60. package/packages/opencode/agent/code-developer.md +11 -12
  61. package/packages/opencode/agent/quality-assurance.md +1 -1
  62. package/packages/{ampcode/commands/systematic-debugging.md → opencode/command/debug-method.md} +9 -9
  63. package/packages/opencode/command/diff-review.md +78 -0
  64. package/packages/opencode/command/friction/friction.js +348 -122
  65. package/packages/opencode/command/optimize.md +45 -4
  66. package/packages/opencode/command/refactor.md +33 -1
  67. package/packages/opencode/command/remember.md +107 -46
  68. package/packages/opencode/command/security.md +28 -1
  69. package/packages/opencode/command/stash.md +7 -0
  70. package/packages/{ampcode/commands/test-driven-development.md → opencode/command/tdd-flow.md} +2 -2
  71. package/packages/opencode/command/test-generate.md +64 -15
  72. package/packages/{ampcode/commands/testing-anti-patterns.md → opencode/command/test-traps.md} +77 -3
  73. package/packages/{ampcode/commands/root-cause-tracing.md → opencode/command/trace-back.md} +3 -3
  74. package/packages/opencode/command/{verification-before-completion.md → verify-done.md} +3 -3
  75. package/packages/opencode/opencode.jsonc +13 -37
  76. package/packages/subagentic-manual.md +55 -51
  77. package/packages/ampcode/commands/code-review.md +0 -107
  78. package/packages/ampcode/commands/condition-based-waiting.md +0 -122
  79. package/packages/ampcode/commands/debug.md +0 -20
  80. package/packages/ampcode/commands/explain.md +0 -18
  81. package/packages/ampcode/commands/friction.md +0 -139
  82. package/packages/ampcode/commands/git-commit.md +0 -14
  83. package/packages/ampcode/commands/review.md +0 -18
  84. package/packages/claude/commands/debug.md +0 -20
  85. package/packages/claude/commands/explain.md +0 -18
  86. package/packages/claude/commands/friction.md +0 -139
  87. package/packages/claude/commands/git-commit.md +0 -14
  88. package/packages/claude/commands/review.md +0 -18
  89. package/packages/claude/skills/code-review/SKILL.md +0 -107
  90. package/packages/claude/skills/code-review/code-reviewer.md +0 -146
  91. package/packages/claude/skills/condition-based-waiting/SKILL.md +0 -122
  92. package/packages/droid/commands/code-review.md +0 -107
  93. package/packages/droid/commands/condition-based-waiting.md +0 -122
  94. package/packages/droid/commands/debug.md +0 -20
  95. package/packages/droid/commands/explain.md +0 -18
  96. package/packages/droid/commands/friction.md +0 -139
  97. package/packages/droid/commands/git-commit.md +0 -14
  98. package/packages/droid/commands/review.md +0 -18
  99. package/packages/opencode/command/code-review.md +0 -107
  100. package/packages/opencode/command/condition-based-waiting.md +0 -122
  101. package/packages/opencode/command/debug.md +0 -20
  102. package/packages/opencode/command/explain.md +0 -18
  103. package/packages/opencode/command/friction.md +0 -139
  104. package/packages/opencode/command/git-commit.md +0 -14
  105. package/packages/opencode/command/review.md +0 -18
  106. /package/packages/ampcode/commands/{condition-based-waiting → test-traps}/example.ts +0 -0
  107. /package/packages/ampcode/commands/{root-cause-tracing → trace-back}/find-polluter.sh +0 -0
  108. /package/packages/claude/skills/{condition-based-waiting → test-traps}/example.ts +0 -0
  109. /package/packages/claude/skills/{root-cause-tracing → trace-back}/find-polluter.sh +0 -0
  110. /package/packages/droid/commands/{condition-based-waiting → test-traps}/example.ts +0 -0
  111. /package/packages/droid/commands/{root-cause-tracing → trace-back}/find-polluter.sh +0 -0
  112. /package/packages/opencode/command/{condition-based-waiting → test-traps}/example.ts +0 -0
  113. /package/packages/opencode/command/{root-cause-tracing → trace-back}/find-polluter.sh +0 -0
package/CHANGELOG.md CHANGED
@@ -17,6 +17,57 @@ and this project adheres to [Semantic Versioning](https://semver.org/spec/v2.0.0
17
17
 
18
18
  ---
19
19
 
20
+ ## [2.10.0] - 2026-06-16
21
+
22
+ ### Changed
23
+ - **`/friction` collapsed into `/remember`; the hot-memory pipeline is now two commands (`/stash → /remember`), down from three.** `/remember` now runs `friction.js` itself as a best-effort first step before consolidating — so the antigen data is always fresh and friction can't be forgotten. Friction targets the tool's **global sessions root** (all projects, since behavioral patterns are cross-project), resolved from an editable, never-prompt probe list baked into `remember.md` (Claude Code, Droid/Factory, Amp, opencode, plus Codex and Antigravity roots; add your own at the top). A no-sessions miss is surfaced **loudly** and degrades to stash-only — never a silent skip. The standalone `/friction` command was removed across all four packages (the `friction.js` script stays, directly runnable for inspection). Counts: claude commands 9→8, droid/opencode/ampcode 17 each.
24
+ - **`/stash` now nudges toward consolidation.** After saving, it derives the unprocessed backlog (`stash files − .processed manifest entries`) and, at ≥5, emits a one-line prompt to run `/remember`. No counter is stored — the count is derived from ground truth, and running `/remember` clears it. The nudge is informational; `/remember` never runs automatically.
25
+
26
+ ### Fixed
27
+ - **`friction.js` no longer crashes on a single malformed JSONL line.** The four mirrored copies parsed session logs and `friction_raw.jsonl` with bare `.map(line => JSON.parse(line))` — one corrupt line aborted the whole run. A new `parseJsonl(raw, source)` helper skips bad lines with a one-line stderr warning (line number + source) and keeps the good records; the whole-file `friction_analysis.json` read is now wrapped in a try/catch that reports the file and bails with exit 1 instead of throwing. Mirrored identically (modulo `.claude`/`.factory`/`.opencode`/`.amp` branding) across all four tool packages.
28
+ - **`packages/subagentic-manual.md` restored after a range-sed corrupted ~310 lines.** A `sed '/start/,/end/{s/.../...}'` earlier this session had `test-generate$` as the end pattern; the range matched far past its intended scope (later occurrences in tree diagrams and category bullets), overwriting the entire tail of the document — Subagents reference, Commands reference, Hot Memory, Usage Patterns, Platform Architecture, Frontmatter Architecture, Contributing — with ~310 duplicate copies of the "Simple Commands" bullet. Restored from the pre-corruption snapshot and re-applied all the renames + count updates that should have happened cleanly. Same fix mirrored to `agentic-toolkit/ai/subagentic/subagentic-manual.md`.
29
+
30
+ ### Changed
31
+ - **Root `package.json` `engines.node` bumped `>=14.0.0` → `>=18.0.0`.** Node 14 has been EOL since 2023-04; the floor now matches the bundled `live-canvas-channel` plugin (`>=18`) and sits well under what CI publishes on (Node 22). As a minimum it excludes no one currently on a supported runtime.
32
+ - **`/review` renamed to `/diff-review` across all four tool packages.** Avoids the name collision with the Anthropic-official `code-review` plugin (which also ships a skill named `review` that operates on PRs). `/diff-review` is more accurate to what the command does — it operates on a diff (staged, working tree, branch range, or against a ref), not on a remote PR. Mirrored into `~/.claude/commands/` and all docs/agents/`opencode.jsonc` references swept.
33
+ - **`/diff-review` absorbed `/code-review`; collapsed to a single command across all four tool packages.** `/diff-review` now accepts a file, a branch (`/diff-review main` diffs `merge-base(main, HEAD)..HEAD` — the common "review my branch before merging" path), or an explicit range (`main..HEAD`). It bakes in the user's standing review focus: bugs needing a fix, dead code, loose ends (added TODO/FIXME, swallowed errors, stubs, abandoned flags), correctness, security, performance, maintainability.
34
+ - **`/diff-review` and `/security` now verify findings, selectively auto-fix, and stop to ask only when needed.** After listing findings, each cited `file:line` is re-grounded in context (and `git grep`'d for dead-code claims) and marked confirmed / false-positive / uncertain. Confirmed + unambiguous + no-contract-change fixes apply directly; the changed region is re-read after the edit. HITL gates fire only for: uncertain findings, multiple reasonable fix shapes, downstream-affecting changes (signatures / response shape / schema / public symbol removal), security primitives (auth / crypto / session / token), or "dead code" that looks intentionally kept. `/diff-review` ends with a one-line **Ready to merge? Yes / No / With fixes** verdict.
35
+ - **Skills renamed to short 2-word slugs (round 2).** `testing-anti-patterns` → **`test-traps`** (now includes timing/polling as AP6 after the fold). `test-driven-development` → **`tdd-flow`** (slug short, H1 short, `TDD` prose preserved as the industry term). `verification-before-completion` → **`verify-done`** (rhythmically mirrors `test-first`). Round 2 paired with round 1 (`systematic-debugging` → `debug-method`, `root-cause-tracing` → `trace-back`) gives a scannable cluster: `tdd-flow / test-generate / test-traps` and `debug-method / trace-back / verify-done`. All references swept across docs, agent files, opencode.jsonc, and the debug-method skill's cross-refs.
36
+ - **Skills renamed to short 2-word slugs (round 1).** `systematic-debugging` → **`debug-method`** (the 4-phase framework with its 4 pressure-test scenarios + creation log preserved). `root-cause-tracing` → **`trace-back`** (the backward-tracing technique with its `find-polluter.sh` bisection helper preserved). Names are shorter, cluster alphabetically under `debug-`, and the "method vs technique" split is now obvious at a glance. All references swept across docs, agent files, and the debug-method skill's internal cross-refs to trace-back.
37
+ - **`/test-generate` rewritten as a generate-and-verify loop, not just a generator.** New flow: discover the existing test framework (refuses to add a new runner) → mirror nearby tests for style/fixtures → generate happy / edge / error cases → **run the new tests** with the project's real test command → **verify each test bites** (mentally swap a broken impl — does the assertion catch it?). Superficial tests (`expect(true).toBe(true)`, mock-asserting-itself, setup-masked passes) count as a failure to ship. Same claim → verify → report shape as `/diff-review` and `/security`. HITL gates: a meaningful test would require a non-obvious design change in production code, ambiguous existing test patterns, or a mock style the project doesn't currently use.
38
+ - **`/optimize` now verifies bottleneck claims before optimizing.** Each cited `file:line` must have at least one of: a profile / benchmark / log line showing call frequency or duration, an obvious hot loop / per-request handler, or user-provided evidence. Unverified claims are marked **uncertain — don't optimize on speculation**. Auto-fixes only when confirmed + unambiguous + no behavior/API change. HITL gates: uncertain (no profile), multiple reasonable shapes (cache vs precompute vs batch vs paginate vs index), public-API / response / schema changes, correctness-for-speed trades, or concurrency primitives.
39
+ - **`/refactor` now runs the tests after the edit.** The "existing tests must pass" constraint was load-bearing but unverified — `/refactor` now detects the project's test command (`package.json` scripts, `pytest`, `go test`, `cargo test`, `Makefile`), runs it scoped to the affected area when possible, and reports pass/fail. If tests fail it **stops and asks** with three options (revert / patch the refactor / update the test with reasoning) rather than auto-reverting (destroys work) or pushing forward (breaks the invariant). Also stops on scope creep and public-API-boundary changes.
40
+ - **`condition-based-waiting` folded into `test-traps` as Anti-Pattern 6: Timeout-Based Waiting.** The two skills covered the same domain (test quality) but only one auto-triggered; folding promotes the timing/polling guidance to auto-trigger coverage. The `example.ts` helper (domain-specific `waitForEvent` / `waitForEventCount` / `waitForEventMatch`) moves with it and is referenced from AP6. Counts: claude skills 10 → 9; droid/opencode/ampcode commands 19 → 18.
41
+ - **CI:** the publish workflow now polls the npm registry for ~2 min (was ~15s; `--prefer-online` skips npm's view cache) and accepts an `exit 0` publish even if the registry hasn't reflected it yet, so a successful-but-slow-to-reflect publish no longer reports a false failure.
42
+ - **`publish.yml` is now manual-only (`workflow_dispatch`) — npm OIDC trusted publishing with provenance, idempotent, and verifies the registry end-state.**
43
+ - **`publish.yml` install step `npm ci` → `npm install`.** This toolkit is dependency-free (no `package-lock.json`), so `npm ci` failed with `EUSAGE`; `npm install` is a fast no-op here and still works if deps are ever added. Removed the superseded manual `scripts/publish.sh` (NPM_TOKEN-via-`pass` flow) — publishing now goes solely through the `publish.yml` GitHub Actions workflow.
44
+
45
+ ### Removed
46
+ - `/code-review` (was: workflow ceremony about *when* to request a review, mostly overlapping `/diff-review`'s purpose). Use `/diff-review` instead — `/diff-review main` for branch-vs-main, `/diff-review` with no args for staged/working-tree.
47
+ - **`/debug`** — was a thin 17-line echo of the `systematic-debugging` skill. The skill (now `debug-method`) carries the real workflow with its pressure-test scenarios; the command added nothing.
48
+ - **`/explain`** — was 11 lines of "explain this code" with no real constraints or workflow. The model does this naturally from a plain prompt.
49
+ - **`/git-commit`** — Claude Code has built-in commit handling and the other three tools don't need a thin wrapper around `git diff --staged` + a templated message either. Use natural-language prompts instead.
50
+
51
+ ---
52
+
53
+ ## [2.9.0] - 2026-05-26
54
+
55
+ Redesign of the `/friction` → `/remember` memory pipeline so friction stops poisoning hot memory and antigens come from what the user actually said. Applied identically across all four tool packages (claude, opencode, ampcode, droid).
56
+
57
+ ### Changed
58
+ - **`/friction` now seeds antigens from observed user reactions, not machine proxies.** Antigen candidates are anchored only on real user reactions (corrections, curses, interrupts); inferred signals survive only as corroborating severity. Clustering is by what the user *said* (content/phrase overlap) instead of `(signal, tool_pattern)`, and recurrence × severity drives a `suggested_artifact` — only patterns recurring across 5+ sessions are meant to load into hot memory. On a 253-session corpus this took false hot preferences from 15 → 0.
59
+ - **`/remember` rewritten to consolidate from friction's short quotes, never raw logs.** It classifies each reaction's target (agent vs. self), drops self-corrections, semantically merges paraphrases that lexical clustering left split, and tiers antigens by recurrence. The generated `MEMORY.md` section is renamed `Preferences` → `Antigens` (High loads hot / Medium recorded / Low = episode).
60
+
61
+ ### Fixed
62
+ - **Terminal pastes are no longer mistaken for friction.** Pasted SSH/shell dumps (prompt lines like `> sudo …` and `root@host:~#`, command output) were captured as user reactions, polluting antigen keywords with shell vocabulary (`postconf`, `qemu`). They are now detected and excluded from both signal detection and keyword extraction, while genuine short corrections that merely mention a command are preserved.
63
+ - **Profanity only counts when it's aimed at the agent.** Narrative/rhetorical curses ("does anyone search any shit?", a pasted reddit story) no longer raise a `user_curse` signal; a curse is kept only in a short reaction turn or when an agent-directed word sits next to it.
64
+ - **Self-corrections are now surfaced to the consolidation step.** The `self_suspect` hint ("wrong project", "nevermind") is propagated from candidate to cluster and rendered in `antigen_review.md`, so `/remember` is told to confirm agent-vs-self target before treating a cluster as an antigen.
65
+
66
+ ### Removed
67
+ - Dead scaffolding left over from the redesign in `friction.js`: the unused `overlap()` helper, the `MIN_KW`/`MIN_INTER` constants, the unread `selfCount` counter (superseded by the `anySelf` flag), and the always-empty `top_files` field with its unreachable renderer block.
68
+
69
+ ---
70
+
20
71
  ## [2.8.3] - 2026-05-24
21
72
 
22
73
  ### Security
package/README.md CHANGED
@@ -9,7 +9,7 @@
9
9
  ╚══════╝╚═╝ ╚═╝ ╚══════╝╚═╝ ╚═╝ ╚═════╝ ╚══════╝╚═╝ ╚═══╝ ╚═╝ ╚══════╝
10
10
  ```
11
11
 
12
- **AI development toolkit with 11 specialized agents and 23 commands per tool**
12
+ **AI development toolkit with 11 specialized agents and 17 commands per tool**
13
13
 
14
14
  <p align="center">
15
15
  <img src="https://img.shields.io/github/package-json/v/hamr0/liteagents?label=version&color=2a4f8c" alt="version (auto from package.json)">
@@ -45,10 +45,10 @@ liteagents
45
45
 
46
46
  ### Supported Tools
47
47
 
48
- - **Claude Code** - 11 subagents + 11 skills + 12 commands (+ optional live-canvas channel plugin)
49
- - **Opencode** - 11 agent references + 23 commands
50
- - **Ampcode** - 11 subagents + 23 commands
51
- - **Droid** - 11 agent references + 23 commands
48
+ - **Claude Code** - 11 subagents + 9 skills + 8 commands (+ optional live-canvas channel plugin)
49
+ - **Opencode** - 11 agent references + 17 commands
50
+ - **Ampcode** - 11 subagents + 17 commands
51
+ - **Droid** - 11 agent references + 17 commands
52
52
 
53
53
  **Key Difference:**
54
54
  - **Claude Code**: Full subagent system with orchestrator + skills (auto-triggering)
@@ -61,30 +61,29 @@ liteagents
61
61
  @orchestrator help
62
62
  @1-create-prd Create a PRD for a task management app
63
63
  /brainstorming Explore authentication approaches
64
- /test-driven-development Implement user login
64
+ /tdd-flow Implement user login
65
65
 
66
66
  # Opencode/Ampcode/Droid examples
67
67
  /1-create-prd Create a PRD for a task management app
68
68
  /brainstorming Explore authentication approaches
69
- /test-driven-development Implement user login
69
+ /tdd-flow Implement user login
70
70
  ```
71
71
 
72
72
  ---
73
73
 
74
74
  ## Hot Memory — project-local learning from your own sessions
75
75
 
76
- Liteagents ships a three-command pipeline that turns Claude Code's session logs into project-local memory. No databases, no external services, just markdown files the assistant reads via `@MEMORY.md`.
76
+ Liteagents ships a two-command pipeline that turns Claude Code's session logs into project-local memory. No databases, no external services, just markdown files the assistant reads via `@MEMORY.md`.
77
77
 
78
78
  ```
79
- /stash → /friction → /remember
80
- capture analyze consolidate
79
+ /stash → /remember
80
+ capture analyze + consolidate
81
81
  ```
82
82
 
83
- - **`/stash`** — snapshot the current session's context before compaction or handoff
84
- - **`/friction`** — mine JSONL session logs for frustration signals, failed flows, and abandonment patterns; cluster them into antigen candidates per project
85
- - **`/remember`** — consolidate stashes + friction antigens into `.claude/memory/MEMORY.md`; auto-injected into `CLAUDE.md` via `@MEMORY.md` so every future session in the project benefits
83
+ - **`/stash`** — snapshot the current session's context before compaction or handoff; nudges you to consolidate once a few stashes pile up
84
+ - **`/remember`** — runs friction analysis automatically (mining JSONL session logs across *all* your projects for frustration signals, failed flows, and abandonment patterns, clustered into antigen candidates), then consolidates stashes + friction antigens into `.claude/memory/MEMORY.md`; auto-injected into `CLAUDE.md` via `@MEMORY.md` so every future session benefits
86
85
 
87
- What you get is a memory that *learns from your own mistakes and interventions*, grows quietly in your repo, and works anywhere Claude Code runs. The friction analyzer alone scans all your projects and gives you a per-repo reliability verdict:
86
+ What you get is a memory that *learns from your own mistakes and interventions*, grows quietly in your repo, and works anywhere Claude Code runs. The friction pass inside `/remember` scans all your projects and gives you a per-repo reliability verdict:
88
87
 
89
88
  ```
90
89
  Per-Project:
@@ -98,7 +97,7 @@ BEST: web-client/0202-2121-8d8608e1 peak=0 turns=4
98
97
  Verdict: USEFUL Intervention predictability: 93%
99
98
  ```
100
99
 
101
- Results land in `.claude/friction/antigen_review.md` with projects, error patterns, and offending tool sequences called out per cluster — so `/remember` can pick them up and encode them as rules the next session sees.
100
+ Results land in `.claude/friction/antigen_review.md` with projects, error patterns, and offending tool sequences called out per cluster — which `/remember` then encodes as rules the next session sees.
102
101
 
103
102
  > This is the thing in liteagents that nothing else ships. Normal skill bundles give you instructions. The hot-memory pipeline gives you instructions the assistant wrote for itself, from your own logs.
104
103
 
@@ -123,14 +122,14 @@ Results land in `.claude/friction/antigen_review.md` with projects, error patter
123
122
  - **system-architect** - System design, technology selection, API design, scalability planning
124
123
  - **ui-designer** - UI/UX design, wireframes, prototypes, accessibility, design systems
125
124
 
126
- ### 23 Commands/Skills
125
+ ### 18 Commands/Skills
127
126
 
128
127
  **Auto-Triggering Skills (3)** - Claude Code only:
129
- - **test-driven-development** - Write test first, watch fail, minimal passing code
130
- - **testing-anti-patterns** - Prevent mocking anti-patterns
131
- - **verification-before-completion** - Verify before claiming done
128
+ - **tdd-flow** - Write test first, watch fail, minimal passing code
129
+ - **test-traps** - Prevent mocking anti-patterns
130
+ - **verify-done** - Verify before claiming done
132
131
 
133
- **Manual Skills/Commands (20):**
132
+ **Manual Skills/Commands (15):**
134
133
 
135
134
  *Hot Memory Pipeline (3)* — see the [Hot Memory](#hot-memory--project-local-learning-from-your-own-sessions) section above for the full walkthrough:
136
135
  - **stash** - Snapshot session context to `.claude/stash/` before compaction, handoff, or ending complex work
@@ -142,19 +141,14 @@ Results land in `.claude/friction/antigen_review.md` with projects, error patter
142
141
 
143
142
  *Workflow & analysis*:
144
143
  - **brainstorming** - Structured brainstorming sessions
145
- - **code-review** - Implementation review against requirements
146
- - **condition-based-waiting** - Replace timeouts with condition polling
147
144
  - **docs-builder** - Project documentation generation
148
- - **root-cause-tracing** - Trace bugs backward through call stack
145
+ - **trace-back** - Trace bugs backward through call stack
149
146
  - **skill-creator** - Guide for creating new skills
150
- - **systematic-debugging** - Four-phase debugging framework
151
- - **debug** - Systematic issue investigation
152
- - **explain** - Explain code for newcomers
153
- - **git-commit** - Intelligent commit creation
147
+ - **debug-method** - Four-phase debugging framework
154
148
  - **optimize** - Performance analysis
155
149
  - **refactor** - Safe refactoring with behavior preservation
156
- - **review** - Comprehensive code review
157
- - **security** - Vulnerability scanning
150
+ - **diff-review** - Review a file, branch, or range; verifies findings, fixes confirmed/unambiguous ones, asks on ambiguous or downstream-affecting ones
151
+ - **security** - Vulnerability scan; same verify→fix→ask flow as `/diff-review`
158
152
  - **ship** - Pre-deployment checklist
159
153
  - **test-generate** - Generate test suites
160
154
 
@@ -186,8 +180,8 @@ Results land in `.claude/friction/antigen_review.md` with projects, error patter
186
180
  **Code Quality:**
187
181
  ```
188
182
  @quality-assurance Review this PR before merge
189
- /code-review Check security and performance
190
- /systematic-debugging Investigate this race condition
183
+ /diff-review main # review branch vs main, fixes confirmed issues, asks on ambiguous ones
184
+ /debug-method Investigate this race condition
191
185
  ```
192
186
 
193
187
  **Architecture & Design:**
package/installer/cli.js CHANGED
@@ -15,7 +15,11 @@ const path = require('path');
15
15
  const readline = require('readline');
16
16
 
17
17
  // Single source of truth for version; UPDATE_VERSION.sh bumps only package.json.
18
- const PACKAGE_VERSION = require('../package.json').version;
18
+ const PACKAGE_JSON = require('../package.json');
19
+ const PACKAGE_VERSION = PACKAGE_JSON.version;
20
+ // Banner counts derived from the description field — same source as README.
21
+ const AGENT_COUNT = (PACKAGE_JSON.description.match(/(\d+)\s+specialized agents/) || [, '11'])[1];
22
+ const COMMAND_COUNT = (PACKAGE_JSON.description.match(/(\d+)\s+commands/) || [, '18'])[1];
19
23
 
20
24
  // ANSI color codes
21
25
  const colors = {
@@ -459,7 +463,7 @@ ${colors.bright}${colors.cyan}██╔══██║██║ ██║█
459
463
  ${colors.bright}${colors.cyan}██║ ██║╚██████╔╝███████╗██║ ╚████║ ██║ ██║╚██████╗ ██║ ██╗██║ ██║${colors.reset}
460
464
  ${colors.bright}${colors.cyan}╚═╝ ╚═╝ ╚═════╝ ╚══════╝╚═╝ ╚═══╝ ╚═╝ ╚═╝ ╚═════╝ ╚═╝ ╚═╝╚═╝ ╚═╝${colors.reset}
461
465
 
462
- ${colors.bright}v${PACKAGE_VERSION} | 11 agents + 23 commands per tool${colors.reset}
466
+ ${colors.bright}v${PACKAGE_VERSION} | ${AGENT_COUNT} agents + ${COMMAND_COUNT} commands per tool${colors.reset}
463
467
  `);
464
468
  }
465
469
 
package/package.json CHANGED
@@ -1,7 +1,7 @@
1
1
  {
2
2
  "name": "liteagents",
3
- "version": "2.8.3",
4
- "description": "AI development toolkit with 11 specialized agents and 23 commands including live-canvas UI design with click-to-annotate feedback. Simple one-question installer for Claude, Opencode, Ampcode, and Droid.",
3
+ "version": "2.10.0",
4
+ "description": "AI development toolkit with 11 specialized agents and 17 commands including live-canvas UI design with click-to-annotate feedback. Simple one-question installer for Claude, Opencode, Ampcode, and Droid.",
5
5
  "main": "index.js",
6
6
  "bin": {
7
7
  "liteagents": "./installer/cli.js",
@@ -59,7 +59,7 @@
59
59
  "access": "public"
60
60
  },
61
61
  "engines": {
62
- "node": ">=14.0.0"
62
+ "node": ">=18.0.0"
63
63
  },
64
64
  "files": [
65
65
  "installer/",
@@ -22,38 +22,32 @@ These subagents are available when using Ampcode CLI.
22
22
  | system-architect | Architect | Use for system design, architecture documents, technology selection, API design, and infrastructure planning |
23
23
  | ui-designer | UX Expert | Use for UI/UX design, wireframes, prototypes, front-end specifications, and user experience optimization |
24
24
 
25
- ### Skills (10 total)
25
+ ### Skills (9 total)
26
26
 
27
27
  | ID | Description | Usage | Auto |
28
28
  |---|---|---|---|
29
29
  | brainstorming | Refines rough ideas into fully-formed designs through collaborative questioning | /brainstorming <session-type> <topic> | false |
30
- | code-review | Reviews implementation against plan or requirements before proceeding | /code-review <review-scope> <focus-areas> | false |
31
- | condition-based-waiting | Replaces arbitrary timeouts with condition polling to wait for actual state changes | /condition-based-waiting <condition-type> <timeout-specs> | false |
32
30
  | docs-builder | Create comprehensive project documentation with structured /docs hierarchy | /docs-builder | false |
33
- | root-cause-tracing | Systematically traces bugs backward through call stack to identify source | /root-cause-tracing <issue-description> | false |
31
+ | trace-back | Systematically traces bugs backward through call stack to identify source | /trace-back <issue-description> | false |
34
32
  | skill-creator | Guide for creating effective skills and extending Claude capabilities | /skill-creator <skill-type> <skill-description> | false |
35
- | systematic-debugging | Four-phase debugging framework - investigate root cause before any fixes | /systematic-debugging <bug-or-error-description> | false |
36
- | test-driven-development | Write test first, watch it fail, write minimal code to pass | /test-driven-development <feature-or-behavior-to-test> | true |
37
- | testing-anti-patterns | Prevents testing mock behavior and production pollution with test-only methods | /testing-anti-patterns <testing-scenario> | true |
38
- | verification-before-completion | Requires running verification commands before making any success claims | /verification-before-completion <work-to-verify> | true |
33
+ | debug-method | Four-phase debugging framework - investigate root cause before any fixes | /debug-method <bug-or-error-description> | false |
34
+ | tdd-flow | Write test first, watch it fail, write minimal code to pass | /tdd-flow <feature-or-behavior-to-test> | true |
35
+ | test-traps | Prevents testing mock behavior and production pollution with test-only methods | /test-traps <testing-scenario> | true |
36
+ | verify-done | Requires running verification commands before making any success claims | /verify-done <work-to-verify> | true |
39
37
 
40
- ### Commands (12 total)
38
+ ### Commands (8 total)
41
39
 
42
40
  | ID | Description | Usage |
43
41
  |---|---|---|
44
- | debug | Debug an issue systematically using structured investigation techniques | /debug <issue-description> |
45
- | explain | Explain code for someone new to the codebase | /explain <code-section> |
46
- | friction | Analyze session logs for failure patterns and behavioral signals | /friction <sessions-path> |
47
- | git-commit | Analyze changes and create intelligent git commits | /git-commit |
48
42
  | live-canvas | Design UI variations and collect click-to-annotate feedback from the browser (batch mode only on Amp) | /live-canvas |
49
43
  | optimize | Analyze and optimize performance issues | /optimize <target-area> |
50
44
  | refactor | Refactor code while maintaining behavior and tests | /refactor <code-section> |
51
45
  | remember | Consolidate stashes + friction into project memory | /remember |
52
- | review | Comprehensive code review including quality, tests, and architecture | /review |
46
+ | diff-review | Comprehensive code review including quality, tests, and architecture | /diff-review |
53
47
  | security | Security vulnerability scan and analysis | /security |
54
48
  | ship | Pre-deployment verification checklist | /ship |
55
49
  | stash | Save session context for compaction recovery or handoffs | /stash ["optional-name"] |
56
- | test-generate | Generate comprehensive test suites for existing code | /test-generate <code-section> |
50
+ | test-generate | Generate tests, run them, verify each one actually exercises the code | /test-generate <file> |
57
51
 
58
52
  All resources are auto-discovered from frontmatter in their respective directories:
59
53
  - **Agents**: `./agents/*.md`
@@ -41,7 +41,7 @@ digraph CodeDeveloper {
41
41
  context_discovery [label="Context Discovery\n(search related code,\ndeps, usages)", fillcolor=lightyellow];
42
42
 
43
43
  // Debug path
44
- use_debug [label="Use /systematic-debugging\nor /root-cause-tracing"];
44
+ use_debug [label="Use /debug-method\nor /trace-back"];
45
45
 
46
46
  // Refactor path
47
47
  use_refactor [label="Use /refactor"];
@@ -54,12 +54,12 @@ digraph CodeDeveloper {
54
54
 
55
55
  // Conditional testing
56
56
  tdd_needed [label="TDD specified\nor tests needed?", shape=diamond];
57
- use_tdd [label="Use /test-driven-development\nor /test-generate"];
57
+ use_tdd [label="Use /tdd-flow\nor /test-generate"];
58
58
 
59
59
  // Validation
60
60
  run_validations [label="Run validations\n(lint, build, tests)"];
61
61
  validations_pass [label="Pass?", shape=diamond];
62
- fix_issues [label="Fix issues\n(use /debug if needed)"];
62
+ fix_issues [label="Fix issues\n(use /debug-method if needed)"];
63
63
  failure_count [label="3+ failures?", shape=diamond];
64
64
 
65
65
  // Security check
@@ -75,8 +75,8 @@ digraph CodeDeveloper {
75
75
  regression_fixable [label="Fixable?", shape=diamond];
76
76
 
77
77
  // Review and complete
78
- code_review [label="Run /code-review"];
79
- verification [label="Run /verification-before-completion", fillcolor=orange];
78
+ code_review [label="Run /diff-review"];
79
+ verification [label="Run /verify-done", fillcolor=orange];
80
80
 
81
81
  // Story-specific
82
82
  update_story [label="Update story\n(checkbox, changelog)"];
@@ -184,15 +184,14 @@ All require `*` prefix. Invocation commands in table above. Additional:
184
184
 
185
185
  | Situation | Delegate To |
186
186
  |-----------|-------------|
187
- | Bug encountered | `/systematic-debugging` first, then `/debug` |
188
- | Error deep in stack | `/root-cause-tracing` |
187
+ | Bug encountered | `/debug-method` (use `/trace-back` when the error is deep in the stack) |
188
+ | Error deep in stack | `/trace-back` |
189
189
  | Refactoring code | `/refactor` |
190
- | Need tests (when required) | `/test-generate` or `/test-driven-development` |
191
- | Writing any test | `/testing-anti-patterns` (avoid mocks, production pollution) |
192
- | Before completion | `/verification-before-completion` |
190
+ | Need tests (when required) | `/test-generate` or `/tdd-flow` |
191
+ | Writing any test | `/test-traps` (avoid mocks, production pollution) |
192
+ | Before completion | `/verify-done` |
193
193
  | After code changes | `/security` |
194
- | Task complete (vs plan) | `/code-review` (checks against requirements/plan) |
195
- | General code review | `/review` (comprehensive quality check) |
194
+ | Task complete / general review | `/diff-review` (diffs branch or staged changes, verifies, fixes confirmed issues, asks on ambiguous ones) |
196
195
  | Performance issues | `/optimize` |
197
196
 
198
197
  You are an autonomous implementation specialist. Execute with precision, delegate appropriately, and communicate clearly when you need guidance or encounter blockers.
@@ -65,7 +65,7 @@ Before any analysis, read (if exists):
65
65
 
66
66
  ## Slash Commands Available
67
67
 
68
- Use these during analysis: `/code-review`, `/security`, `/debug`, `/review`, `/verification-before-completion`
68
+ Use these during analysis: `/diff-review`, `/security`, `/verify-done`
69
69
 
70
70
  ## Analysis Areas
71
71
 
@@ -1,11 +1,11 @@
1
1
  ---
2
- name: systematic-debugging
2
+ name: debug-method
3
3
  description: Use when encountering any bug, test failure, or unexpected behavior, before proposing fixes - four-phase framework (root cause investigation, pattern analysis, hypothesis testing, implementation) that ensures understanding before attempting solutions
4
- usage: /systematic-debugging <bug-or-error-description>
4
+ usage: /debug-method <bug-or-error-description>
5
5
  auto_trigger: false
6
6
  ---
7
7
 
8
- # Systematic Debugging
8
+ # Debug Method
9
9
 
10
10
  ## Overview
11
11
 
@@ -113,7 +113,7 @@ You MUST complete each phase before proceeding to the next.
113
113
 
114
114
  **WHEN error is deep in call stack:**
115
115
 
116
- **REQUIRED SUB-SKILL:** Use root-cause-tracing for backward tracing technique
116
+ **REQUIRED SUB-SKILL:** Use trace-back for backward tracing technique
117
117
 
118
118
  **Quick version:**
119
119
  - Where does bad value originate?
@@ -178,7 +178,7 @@ You MUST complete each phase before proceeding to the next.
178
178
  - Automated test if possible
179
179
  - One-off test script if no framework
180
180
  - MUST have before fixing
181
- - **REQUIRED SUB-SKILL:** Use test-driven-development for writing proper failing tests
181
+ - **REQUIRED SUB-SKILL:** Use tdd-flow for writing proper failing tests
182
182
 
183
183
  2. **Implement Single Fix**
184
184
  - Address the root cause identified
@@ -280,13 +280,13 @@ If systematic investigation reveals issue is truly environmental, timing-depende
280
280
  ## Integration with Other Skills
281
281
 
282
282
  **This skill requires using:**
283
- - **root-cause-tracing** - REQUIRED when error is deep in call stack (see Phase 1, Step 5)
284
- - **test-driven-development** - REQUIRED for creating failing test case (see Phase 4, Step 1)
283
+ - **trace-back** - REQUIRED when error is deep in call stack (see Phase 1, Step 5)
284
+ - **tdd-flow** - REQUIRED for creating failing test case (see Phase 4, Step 1)
285
285
 
286
286
  **Complementary skills:**
287
287
  - **defense-in-depth** - Add validation at multiple layers after finding root cause
288
- - **condition-based-waiting** - Replace arbitrary timeouts identified in Phase 2
289
- - **verification-before-completion** - Verify fix worked before claiming success
288
+ - **test-traps** (Anti-Pattern 6) - Replace arbitrary timeouts identified in Phase 2
289
+ - **verify-done** - Verify fix worked before claiming success
290
290
 
291
291
  ## Real-World Impact
292
292
 
@@ -0,0 +1,78 @@
1
+ ---
2
+ name: diff-review
3
+ description: Review diff [file, branch, or range]
4
+ usage: /diff-review
5
+ argument-hint: [file, branch (e.g. main), range (main..HEAD), or empty]
6
+ allowed-tools: Read, Edit, Grep, Glob, Bash(git diff *), Bash(git log *), Bash(git show *), Bash(git status *), Bash(git grep *), Bash(git rev-parse *), Bash(git merge-base *), Bash(rg *)
7
+ ---
8
+ Review $ARGUMENTS. Interpret in this order:
9
+ 1. **Empty** → staged diff (`git diff --staged`); if empty, working-tree diff
10
+ (`git diff`).
11
+ 2. **A range** like `main..HEAD` or `origin/main...HEAD` → `git diff <range>`.
12
+ 3. **A single ref** (branch / tag / SHA — confirm with `git rev-parse
13
+ --verify`) → diff that ref's merge-base against `HEAD` (i.e. everything on
14
+ the current branch since it diverged: `git diff $(git merge-base <ref>
15
+ HEAD)..HEAD`). This is the common "review my branch before merging" path.
16
+ 4. **A file or directory path** → that target.
17
+ 5. Otherwise → ask.
18
+
19
+ The diff is the subject; widen to surrounding code only as needed to judge a
20
+ hunk. For multi-commit ranges, also skim `git log <range>` to understand
21
+ intent before judging.
22
+
23
+ ## Check For
24
+ - **Bugs needing a fix.** Logic errors, off-by-one, null/undefined paths,
25
+ races, wrong defaults, broken edge cases. Concrete failure modes only — not
26
+ vibes.
27
+ - **Dead code.** Unreferenced functions / vars / imports / params, unreachable
28
+ branches, commented-out blocks, legacy paths the diff just obsoleted.
29
+ `git grep` the symbol before flagging — easy to be wrong.
30
+ - **Loose ends.** TODO / FIXME / XXX added by this diff, half-finished
31
+ branches, silently swallowed errors, stub bodies, mocked-out paths,
32
+ "temporary" names, abandoned feature flags.
33
+ - **Correctness.** Edge cases, error handling, type / contract violations,
34
+ broken invariants.
35
+ - **Security.** OWASP Top 10, auth, data exposure. (`/security` for depth.)
36
+ - **Performance.** N+1, blocking calls in hot paths, unbounded loops, indexes
37
+ the diff actually touches.
38
+ - **Maintainability.** Complexity, naming, duplication — only when material.
39
+
40
+ ## Output Format
41
+ ### 🚨 Critical (blocks merge)
42
+ ### ⚠️ Warnings (should fix)
43
+ ### 💡 Suggestions (nice to have)
44
+
45
+ Each finding: **Location** (`file:line`), **What's wrong**, **Why it matters**,
46
+ **Concrete fix** — not "consider improving".
47
+
48
+ ## After the review — verify, then fix
49
+
50
+ Findings are claims, not facts. Validate before acting; validate again after.
51
+
52
+ **Verify each claim.** Re-read the cited `file:line` in context. For
53
+ dead-code or unused-symbol claims, `git grep` the name across the repo before
54
+ trusting it. Mark each **confirmed**, **false positive** (with reason), or
55
+ **uncertain**.
56
+
57
+ **Fix what's confirmed and unambiguous** — minimal shape, one obvious way, no
58
+ change to a public API / response / caller contract. Apply directly. After
59
+ each edit, re-read the changed region and confirm it does what you intended
60
+ without breaking nearby logic. A fix isn't done until you've grounded it the
61
+ same way you grounded the claim.
62
+
63
+ **Stop and ask** when any of these hold (HITL gates — not all the time, only
64
+ here):
65
+ - the finding is **uncertain** after grounding,
66
+ - the fix has **multiple reasonable shapes** (e.g. delete-vs-keep-behind-flag,
67
+ extract-vs-inline, patch-vs-rewrite) — present options with tradeoffs, not a
68
+ chosen path,
69
+ - it **affects downstream** (signatures, response shape, schema, any caller
70
+ contract) or removes a public/exported symbol, or
71
+ - the "dead code" looks intentionally kept (stub for upcoming work, framework
72
+ hook, documented extension point) — confirm before deleting.
73
+
74
+ Final report: **confirmed-and-fixed** · **confirmed-but-asking** (why +
75
+ options) · **false-positive** (why) · **uncertain** (what's needed to decide).
76
+
77
+ End with a one-line verdict: **Ready to merge? Yes / No / With fixes** — and
78
+ the reason in a sentence.