@rhize/skill-forge 0.17.1 → 0.20.0

This diff represents the content of publicly available package versions that have been released to one of the supported registries. The information contained in this diff is provided for informational purposes only and reflects changes between package versions as they appear in their respective public registries.
package/LICENSE CHANGED
@@ -27,6 +27,7 @@ Pro Modules
27
27
  - src/refine/store.ts
28
28
  - src/refine/patch.ts
29
29
  - src/refine/context.ts
30
+ - src/refine/activation.ts
30
31
  - src/insights/models.ts
31
32
  - src/insights/sources.ts
32
33
  - src/insights/store.ts
package/README.md CHANGED
@@ -2,8 +2,9 @@
2
2
 
3
3
  **The supply-chain gate for agent skills.**
4
4
 
5
- Status: published — `@rhize/skill-forge@0.17.1` on npm (released 2026-09-05),
6
- 0.x beta (Pro features free until 1.0)
5
+ Source version: `@rhize/skill-forge@0.20.0` (2026-09-14), 0.x beta (Pro features free until 1.0).
6
+ Npm publication is a separate tag-driven step; `npm view @rhize/skill-forge version` reports the
7
+ latest published version.
7
8
 
8
9
  > This project is unrelated to the [`skillforge`](https://www.npmjs.com/package/skillforge)
9
10
  > package on npm, which is a Claude Skills *evaluation* framework. `skill-forge` (this package,
@@ -11,16 +12,17 @@ Status: published — `@rhize/skill-forge@0.17.1` on npm (released 2026-09-05),
11
12
 
12
13
  ## Why
13
14
 
14
- Agent skills (Claude Code skills, and the equivalent extension formats emerging in Cursor and
15
- Codex) are now distributed the way npm packages were a decade ago: a public registry, a one-line
16
- install command, and no vetting step in between. skills.sh alone lists 600k+ skills. A single
17
- `npx skills@latest add owner/name` drops arbitrary third-party instructions, scripts, and tool
18
- bindings directly into a working agent's context and file system — with the same trust model as
19
- copy-pasting a shell script from a stranger.
15
+ Agent skills can include third-party instructions and executable resources. Existing marketplaces
16
+ already provide safeguards: [skills.sh offers partner security audits and installation risk
17
+ information](https://vercel.com/changelog/automated-security-audits-now-available-for-skills-sh).
18
+ Skill Forge adds a local review and maintenance workflow around the skills you actually use:
19
+ provenance, overlap review, drift checks, and recoverable project refinements.
20
20
 
21
- skill-forge does not replace that install command; it wraps it. Every candidate skill is routed
22
- through **quarantine → profile → safety scan → overlap analysis → report → an explicit
23
- promote/hold/reject decision**, before it is allowed anywhere near your working skill set.
21
+ Candidates submitted through Skill Forge follow **quarantine → profile → safety scan → overlap
22
+ analysis (when available/configured) → report → an explicit promote/hold/reject decision**.
23
+ This gate applies to the Skill Forge path; it does not prevent direct installs through other tools
24
+ or replace host-managed restrictions. Passing a scan does not prove that a skill is safe.
25
+ See the [product spec](docs/PRODUCT.md) for current competitive positioning and evidence limits.
24
26
 
25
27
  ## Quickstart
26
28
 
@@ -66,7 +68,8 @@ reports every enabled plugin (safety findings, skill count, and — with `--usag
66
68
  `unknown` recommendation). Plugin roots are never promotion targets, the safety scan surface
67
69
  never shrinks, and nothing here changes an `add`/`scan` verdict. `init` asks once whether to
68
70
  enable this; `routine` carries the same flags for cron. See
69
- [docs/commands/audit.md](docs/commands/audit.md).
71
+ [docs/commands/audit.md](docs/commands/audit.md). For a practical maintenance loop and the limits
72
+ of invocation-based advice, see [Plugin governance](docs/plugin-governance.md).
70
73
 
71
74
  `<source>` accepts a `skills.sh` `owner/name` slug, a git URL (`https://...`, `git@...`, or
72
75
  anything ending in `.git`), or a local filesystem path.
@@ -85,6 +88,7 @@ anything ending in `.git`), or a local filesystem path.
85
88
  | `find [query] [options]` | Discover skills via skills.sh and check partner security audits (free) | [docs/commands/find.md](docs/commands/find.md) |
86
89
  | `watch [options]` (Pro) | Drift check across every `SOURCES.md` provenance ledger | [docs/commands/watch.md](docs/commands/watch.md) |
87
90
  | `ingest [options]` / `queue close <id>` (Pro) | Hand the pending-ingestion queue off to a coding agent for the decide/absorb pass | [docs/commands/ingest.md](docs/commands/ingest.md) |
91
+ | `plugins keep/list/forget` | Record explicit host/project keep rationale without suppressing safety findings | [governance additions](docs/governance-improvements.md) |
88
92
  | `refine [options]` (Pro) | Capture, list, review, promote, and roll back project-scope skill overrides | [docs/commands/refine.md](docs/commands/refine.md) |
89
93
  | `insight <subcommand>` (Pro) | Turn sources into evidence-backed capability plans and Jira manifests | [docs/commands/insight.md](docs/commands/insight.md) |
90
94
  | `routine [options]` (Pro) | One scheduled maintenance pass: audit (plugins + usage aware) + drift + registry, cron-friendly | [docs/commands/routine.md](docs/commands/routine.md) |
@@ -154,8 +158,9 @@ including current caveats, in [docs/configuration.md](docs/configuration.md).
154
158
  **Is this related to the `skillforge` npm package?**
155
159
  No. [`skillforge`](https://www.npmjs.com/package/skillforge) is a Claude Skills *evaluation*
156
160
  framework — it tests whether a skill performs well. `skill-forge` (this package, hyphenated) is a
157
- supply-chain security gate for skill *installation* — it decides whether a skill is safe and
158
- non-redundant before it ever runs. Same neighborhood, different problem, name collision only.
161
+ review gate for skill *installation* and ongoing collection maintenance. It reports detected
162
+ risks and potential overlap; it does not certify safety or task effectiveness. The similarly named
163
+ packages are independently maintained.
159
164
 
160
165
  **Does skill-forge replace `npx skills@latest add`?**
161
166
  No, it wraps it. `add` uses the same install mechanisms (skills.sh, git, local copy) but stages
@@ -206,3 +211,26 @@ skill-forge is open-core with a split license (as of v0.2.0):
206
211
 
207
212
  [LICENSE](LICENSE) is the authoritative map of which files fall under which license.
208
213
  Versions up to and including 0.1.0 were published entirely under MIT.
214
+
215
+ ## Governance utility additions (0.19)
216
+
217
+ Use `plugins keep/list/forget` for explicit host/project retention rationale without
218
+ suppressing safety findings. `audit --plugin-inventory <file>` and the same option
219
+ on `routine` import static host evidence; Codex never joins Claude usage counts.
220
+ `audit --compare <full-audit.json>` highlights changes while preserving all findings.
221
+ `refine activate --skill <capture> --base <dir> --as <unique-name> --host claude|codex`
222
+ materializes a captured patch/extend as a new project skill after a staged safety scan.
223
+ Preview with `--dry-run --json`; activation requires confirmation. Verify actual host
224
+ loading in a fresh task before treating the refinement as consumed. Roll back using
225
+ the returned backup ID with `refine rollback <id> --project --yes`.
226
+
227
+ See the [governance operating guide](docs/governance-improvements.md) for scope, schemas, examples and limits.
228
+
229
+ ## Advisory instruction diagnostics (0.20)
230
+
231
+ `audit` reports description, entrypoint and resource measurements separately, with review
232
+ suggestions for broad triggers and resource routing. These suggestions never change safety
233
+ findings, gate decisions or installed skills. Static reports keep actual host loading and
234
+ truncation unknown; rough token estimates are not measured usage. Comparisons show covered
235
+ skills and description/body deltas when both reports contain the new evidence. See
236
+ [the audit command](docs/commands/audit.md) for metric definitions and unavailable-resource handling.