musubix3 0.1.14 → 0.1.15

This diff represents the content of publicly available package versions that have been released to one of the supported registries. The information contained in this diff is provided for informational purposes only and reflects changes between package versions as they appear in their respective public registries.
@@ -3,13 +3,13 @@
3
3
  "owner": { "name": "nahisaho" },
4
4
  "metadata": {
5
5
  "description": "GitHub Copilot CLI specification-driven development skills",
6
- "version": "0.1.14"
6
+ "version": "0.1.15"
7
7
  },
8
8
  "plugins": [
9
9
  {
10
10
  "name": "musubix3",
11
11
  "source": ".",
12
- "version": "0.1.14",
12
+ "version": "0.1.15",
13
13
  "description": "Evidence-driven SDD without duplicating native Copilot capabilities."
14
14
  }
15
15
  ]
@@ -9,6 +9,7 @@ description: "Use as the MANDATORY first Skill for requests to develop, build, c
9
9
  */
10
10
  Mandatory entrypoint: every new natural-language development request is a new change, even in an existing Copilot session; never reuse prior requirements, approvals, TDD, or change evidence unless the user explicitly names the existing change ID and asks to continue it. never start implementation before validating requirements/design; skip only for verified approved artifacts of that explicitly continued change.
11
11
  Never infer approval; show `approval prepare <stage>` and record only its reviewed hash with `approval record <stage> --approver <name> --artifact-sha256 <hash> --confirm`.
12
+ Whenever an AI deliverable is documentation (requirements, design, ADRs, the CHANGE document, or release/quality evidence), run Copilot's native `rubber-duck` review agent on it before that phase's human approval, fixing every issue and re-reviewing until none remain.
12
13
  Follow the user's input language. Use native Copilot planning, editing, research, review, security review and subagents.
13
14
  Record exactly one final invocation outcome with `npx musubix3 workflow-record sdd-change complete --status <status>`; `change-record` separately proves phases.
14
15
  Run `workflow-sanitize <copilot.jsonl> <safe.jsonl>` before review, then
@@ -38,11 +39,10 @@ Persisted monotonic order, not wall-clock time, proves these phase boundaries.
38
39
  2. For a bug where implementation violates an existing requirement, keep that
39
40
  requirement and record that no specification change is needed. Never rewrite
40
41
  a requirement merely to make incorrect behavior appear compliant.
41
- 3. Update design responsibilities/interfaces/constraints/links and ADRs for real decisions.
42
- 4. Run requirements, constitution and design validation. Stop on invalid
43
- artifacts instead of continuing with unapproved assumptions.
44
- 5. Stop before design for explicit current `requirements` approval; stop before
45
- Red/implementation for explicit current `design` approval. Changes re-open approval.
42
+ 3. Run requirements and constitution validation; stop on invalid artifacts
43
+ instead of continuing with unapproved assumptions.
44
+ 4. Run a `rubber-duck` review (Copilot's native review agent) of `requirements.md`, fixing every issue and re-reviewing until none remain, then stop for explicit current `requirements` approval before design. An edit to `requirements.md` re-opens approval and requires re-review.
45
+ 5. Update design responsibilities/interfaces/constraints/links and ADRs, run design validation, then run the same rubber-duck review/fix loop on `design.md`/ADRs before stopping for explicit current `design` approval before Red/implementation. An edit re-opens approval and requires re-review.
46
46
  ## 3. Implement and prove coverage / 実装と網羅性
47
47
  1. For observable behavior changes and defect fixes, write the smallest meaningful
48
48
  test first. Include its `TEST-*` ID in the test name/output and link it to the
@@ -70,8 +70,6 @@ Documentation/prototypes may omit TDD only when policy allows; record the reason
70
70
  A required `formal` check enforces modeled fraction and configured solver.
71
71
  3. Run `gate --changed --json` and `status --json`. Repair failures, dangling
72
72
  links and stale evidence; never weaken requirements or policy to obtain green.
73
- 4. Treat the first otherwise-passing gate as the release candidate. Before any
74
- release operation, prepare/show its exact hash, ask one human approve/reject
75
- question and wait; record only that hash, then rerun gate/status.
73
+ 4. Treat the first otherwise-passing gate as the release candidate. Run a `rubber-duck` review of the release/quality evidence summary and the CHANGE document; fix every reported issue and re-review until zero issues remain. Before any release operation, prepare/show its exact hash, ask one human approve/reject question and wait; record only that hash, then rerun gate/status.
76
74
  5. Complete only when required checks pass and `status.gate.ready` is true;
77
75
  otherwise report blockers. One-phase work must state downstream work.
@@ -25,8 +25,12 @@ completed` exactly once.
25
25
  5. Run `npx musubix3 trace build` then `npx musubix3 trace check`.
26
26
  Do not hand-edit `trace.json`. Ask native review to inspect coupling and
27
27
  coverage; use `sdd-formal-codegraph` for compiler-based impact checks.
28
- 6. Before implementation or Red, run `approval prepare design`, show its exact
28
+ 6. Before requesting human approval, run Copilot's native `rubber-duck`
29
+ review agent on `design.md` and any related ADRs. Fix every reported
30
+ issue, then re-run the review. Repeat fix-then-review until the review
31
+ reports zero remaining issues; only then proceed to step 7.
32
+ 7. Before implementation or Red, run `approval prepare design`, show its exact
29
33
  artifacts/hash, ask one explicit approve/reject question, then wait. On approval
30
34
  only, record that hash with `approval record design --approver <name>
31
35
  --artifact-sha256 <hash> --confirm`; rejection or changes require renewed review.
32
- 7. Hand off responsibilities, interfaces, constraints and IDs to implementation.
36
+ 8. Hand off responsibilities, interfaces, constraints and IDs to implementation.
@@ -71,8 +71,5 @@ completed` exactly once.
71
71
  check all configured claims; offline strict verification fails closed.
72
72
  Ensure the baseline protects CI-required mode, strict OIDC and key binding.
73
73
  Never store a private key or overstate OIDC as proof of arbitrary runner work.
74
- 7. After required non-approval checks pass, run `approval prepare release`, show
75
- its exact hash and residual risks, ask one approve/reject question, then wait.
76
- On approval only record that hash with `approval record release --approver
77
- <name> --artifact-sha256 <hash> --confirm`; rejection stops. Rerun gate/status;
78
- stale approval blocks commit/push/publish/deploy; no resident watcher or REPL.
74
+ 7. Before requesting human approval, run Copilot's native `rubber-duck` review agent on the release evidence summary and the CHANGE document; fix every reported issue, then re-run the review, repeating until it reports zero remaining issues.
75
+ 8. After required non-approval checks pass, run `approval prepare release`, show its exact hash and residual risks, ask one approve/reject question, then wait. On approval only record that hash with `approval record release --approver <name> --artifact-sha256 <hash> --confirm`; rejection stops. Rerun gate/status; stale approval blocks commit/push/publish/deploy; no resident watcher or REPL.
@@ -50,10 +50,15 @@ After the work, run `npx musubix3 workflow-record sdd-requirements complete
50
50
  5. Validate with `npx musubix3 requirements validate <file> --json` and
51
51
  `npx musubix3 constitution validate --json`. Rules use `PRINC-001`, `RULE-001`,
52
52
  a supported `Metric:` and numeric `Limit:`. Validation is not execution evidence.
53
- 6. Validation is not human approval. Before design, run `approval prepare
53
+ 6. Before requesting human approval, run Copilot's native `rubber-duck`
54
+ review agent on `requirements.md`. Fix every reported issue, then re-run
55
+ the review. Repeat fix-then-review until the review reports zero
56
+ remaining issues; only then proceed to step 7. Do not request human
57
+ approval while rubber-duck findings remain open.
58
+ 7. Validation is not human approval. Before design, run `approval prepare
54
59
  requirements`, show its exact artifacts/hash, ask one explicit approve/reject
55
60
  question, then wait. On approval only, record that hash with `approval record
56
61
  requirements --approver <name> --artifact-sha256 <hash> --confirm`; rejection
57
62
  or any intervening artifact change stops and requires renewed review.
58
- 7. Hand off confirmed IDs and unresolved assumptions to `sdd-design`; use native
63
+ 8. Hand off confirmed IDs and unresolved assumptions to `sdd-design`; use native
59
64
  review for semantic completeness, not merely syntactic conformance.
package/CHANGELOG.md CHANGED
@@ -1,5 +1,23 @@
1
1
  # Changelog
2
2
 
3
+ ## 0.1.15 - 2026-09-11
4
+
5
+ Documentation/workflow update: AI-generated documentation deliverables now go
6
+ through a mandatory rubber-duck review/fix loop before human approval.
7
+
8
+ - `sdd-requirements`, `sdd-design`, `sdd-quality`, and `sdd-change` now
9
+ require a `rubber-duck` review of any AI-generated documentation artifact
10
+ (requirements.md, design.md/ADRs, the CHANGE document, and release/quality
11
+ evidence summaries) before the corresponding human approval step
12
+ (`requirements`, `design`, or `release`). Every reported issue must be
13
+ fixed and the artifact re-reviewed; only once the review reports zero
14
+ remaining issues may human review/approval be requested.
15
+ - Documented the same review/fix loop in the Workflow section of
16
+ README.md/README-ja.md.
17
+ - No CLI, schema, or validation behavior changed. Apart from the package
18
+ version metadata bump (`package.json`/`package-lock.json`), this release
19
+ only updates the bundled `.github/skills/sdd-*` guidance and documentation.
20
+
3
21
  ## 0.1.14 - 2026-09-11
4
22
 
5
23
  Fix for one issue (#17) found while running musubix3@0.1.13 against a large
package/README-ja.md CHANGED
@@ -248,17 +248,25 @@ Copilot の提案を記号的検査で制約する構成であり、独自の「
248
248
  要求を確定しません。
249
249
 
250
250
  1. ネイティブ計画・調査で意図と測定可能な受入条件を明確化。
251
- 2. `change-record CHANGE-ID impact` を記録して要求を編集・検証し、
252
- `requirements` checkpoint後、設計前にartifact-boundな人間の
253
- `requirements`承認を明示的に記録。
254
- 3. コンポーネントとADRを更新して `design` checkpointを記録し、
255
- 実装前にartifact-boundな人間の`design`承認を明示的に記録。
251
+ 2. `change-record CHANGE-ID impact` を記録して要求を編集・検証し、要求文書の
252
+ `rubber-duck`レビューを実施して指摘事項をすべて修正(指摘がゼロになるまで
253
+ レビュー・修正を繰り返す)、`requirements` checkpoint後、設計前に
254
+ artifact-boundな人間の`requirements`承認を明示的に記録。
255
+ 3. コンポーネントとADRを更新し、設計文書・ADRの`rubber-duck`レビューを実施して
256
+ 指摘事項をすべて修正(指摘がゼロになるまで繰り返す)、`design` checkpoint
257
+ を記録し、実装前にartifact-boundな人間の`design`承認を明示的に記録。
256
258
  4. 注釈付きテストを作成し、構造化結果を伴う `tdd red` と変更の `red` を記録。
257
259
  5. 最小実装後に `implementation`、成功する `tdd green`、変更の `green` を記録。
258
260
  6. 注釈とグラフを更新して変更影響・網羅性を確認。
259
- 7. 実コマンドで候補品質ゲートを実行。承認以外の必須checkがすべて合格した後、
260
- 人間が`release`承認を記録してgate/statusを再実行し、その後だけ
261
- commit・push・publish・deployへ進む。検証結果や自然言語から承認を推測しない。
261
+ 7. 実コマンドで候補品質ゲートを実行。リリース・品質エビデンス要約の
262
+ `rubber-duck`レビューを実施して指摘事項をすべて修正(指摘がゼロになるまで
263
+ 繰り返す)。承認以外の必須checkがすべて合格した後、人間が`release`承認を
264
+ 記録してgate/statusを再実行し、その後だけcommit・push・publish・deployへ
265
+ 進む。検証結果や自然言語から承認を推測しない。
266
+
267
+ AIが生成した文書成果物(要求、設計、ADR、CHANGE文書、リリース・品質エビデンス
268
+ 要約)は、対応する人間承認を依頼する**前**に、指摘事項がゼロになるまでこの
269
+ `rubber-duck`レビュー・修正ループを実施する。
262
270
 
263
271
  ```sh
264
272
  npx musubix3 requirements validate .musubix/features/example/requirements.md --json
package/README.md CHANGED
@@ -252,11 +252,15 @@ question at a time and waits for the answer; it does not batch questions or
252
252
  finalize requirements while blockers remain.
253
253
 
254
254
  1. Use native planning/research to establish intent and measurable acceptance.
255
- 2. Record `change-record CHANGE-ID impact`, edit/validate requirements, record
256
- the `requirements` checkpoint, then obtain explicit artifact-bound human
257
- `requirements` approval before design.
258
- 3. Design explicit components, record trade-offs in ADRs, record `design`, then
259
- obtain explicit artifact-bound human `design` approval before implementation.
255
+ 2. Record `change-record CHANGE-ID impact`, edit/validate requirements, run a
256
+ `rubber-duck` review of the requirements document and fix every finding
257
+ (repeat review/fix until zero issues remain), record the `requirements`
258
+ checkpoint, then obtain explicit artifact-bound human `requirements`
259
+ approval before design.
260
+ 3. Design explicit components, record trade-offs in ADRs, run a `rubber-duck`
261
+ review of the design document/ADRs and fix every finding (repeat until zero
262
+ issues remain), record `design`, then obtain explicit artifact-bound human
263
+ `design` approval before implementation.
260
264
  4. Write an annotated behavior test, record a structured failing `tdd red`, then
261
265
  record the change `red` checkpoint.
262
266
  5. Implement the minimum change, record `implementation`, run passing `tdd green`,
@@ -266,10 +270,17 @@ finalize requirements while blockers remain.
266
270
  Red-Implementation-Green loop before moving to the next, instead of
267
271
  completing every requirement's Red before any Implementation.
268
272
  6. Add trace annotations, build graphs, inspect impact and fix missing coverage.
269
- 7. Configure real checks and run the candidate `gate --changed`. After every
270
- required non-approval check passes, obtain explicit human `release` approval,
273
+ 7. Configure real checks and run the candidate `gate --changed`. Run a
274
+ `rubber-duck` review of the release/quality evidence summary and fix every
275
+ finding (repeat until zero issues remain). After every required
276
+ non-approval check passes, obtain explicit human `release` approval,
271
277
  rerun the gate/status, and only then commit, push, publish, or deploy.
272
278
 
279
+ Any AI-generated documentation deliverable (requirements, design, ADRs, the
280
+ CHANGE document, or release/quality evidence summaries) goes through this
281
+ `rubber-duck` review/fix loop until zero issues remain **before** the
282
+ corresponding human approval step is requested.
283
+
273
284
  ```sh
274
285
  npx musubix3 requirements validate .musubix/features/example/requirements.md --json
275
286
  npx musubix3 constitution validate --json
@@ -24,7 +24,7 @@ function pathQuery(root, query) {
24
24
  return portable(relative(root, resolve(root, query)));
25
25
  }
26
26
  export function createProgram() {
27
- const program = new Command().name('musubix3').description('Evidence-driven SDD for GitHub Copilot CLI / 根拠に基づく仕様駆動開発').version('0.1.14');
27
+ const program = new Command().name('musubix3').description('Evidence-driven SDD for GitHub Copilot CLI / 根拠に基づく仕様駆動開発').version('0.1.15');
28
28
  program.exitOverride();
29
29
  common(program.command('init').alias('install').description('Install repository skills and SDD artifacts (preserves existing files)'))
30
30
  .option('--dry-run', 'Preview without writing').option('--force', 'Replace bundled, managed paths only')
package/package.json CHANGED
@@ -1,6 +1,6 @@
1
1
  {
2
2
  "name": "musubix3",
3
- "version": "0.1.14",
3
+ "version": "0.1.15",
4
4
  "description": "Evidence-driven specification development skills for GitHub Copilot CLI",
5
5
  "type": "module",
6
6
  "license": "MIT",
package/plugin.json CHANGED
@@ -1,6 +1,6 @@
1
1
  {
2
2
  "name": "musubix3",
3
- "version": "0.1.14",
3
+ "version": "0.1.15",
4
4
  "description": "Evidence-driven SDD skills for GitHub Copilot CLI: requirements, design, traceability, implementation, quality, knowledge, formal checks and codegraph.",
5
5
  "author": { "name": "nahisaho" },
6
6
  "repository": "https://github.com/nahisaho/musubix3",