@claude-flow/cli 3.28.0 → 3.30.0

This diff represents the content of publicly available package versions that have been released to one of the supported registries. The information contained in this diff is provided for informational purposes only and reflects changes between package versions as they appear in their respective public registries.
Files changed (465) hide show
  1. package/.claude/agents/analysis/analyze-code-quality.md +178 -178
  2. package/.claude/agents/analysis/code-analyzer.md +209 -209
  3. package/.claude/agents/analysis/code-review/analyze-code-quality.md +178 -178
  4. package/.claude/agents/architecture/arch-system-design.md +156 -156
  5. package/.claude/agents/architecture/system-design/arch-system-design.md +154 -154
  6. package/.claude/agents/browser/browser-agent.yaml +182 -182
  7. package/.claude/agents/consensus/byzantine-coordinator.md +62 -62
  8. package/.claude/agents/consensus/crdt-synchronizer.md +996 -996
  9. package/.claude/agents/consensus/gossip-coordinator.md +62 -62
  10. package/.claude/agents/consensus/performance-benchmarker.md +850 -850
  11. package/.claude/agents/consensus/quorum-manager.md +822 -822
  12. package/.claude/agents/consensus/raft-manager.md +62 -62
  13. package/.claude/agents/consensus/security-manager.md +621 -621
  14. package/.claude/agents/core/planner.md +374 -374
  15. package/.claude/agents/custom/test-long-runner.md +44 -44
  16. package/.claude/agents/data/data-ml-model.md +444 -444
  17. package/.claude/agents/data/ml/data-ml-model.md +192 -192
  18. package/.claude/agents/development/backend/dev-backend-api.md +141 -141
  19. package/.claude/agents/development/dev-backend-api.md +344 -344
  20. package/.claude/agents/devops/ci-cd/ops-cicd-github.md +163 -163
  21. package/.claude/agents/devops/ops-cicd-github.md +164 -164
  22. package/.claude/agents/documentation/api-docs/docs-api-openapi.md +173 -173
  23. package/.claude/agents/documentation/docs-api-openapi.md +354 -354
  24. package/.claude/agents/flow-nexus/app-store.md +87 -87
  25. package/.claude/agents/flow-nexus/authentication.md +68 -68
  26. package/.claude/agents/flow-nexus/challenges.md +80 -80
  27. package/.claude/agents/flow-nexus/neural-network.md +87 -87
  28. package/.claude/agents/flow-nexus/payments.md +82 -82
  29. package/.claude/agents/flow-nexus/sandbox.md +75 -75
  30. package/.claude/agents/flow-nexus/swarm.md +75 -75
  31. package/.claude/agents/flow-nexus/user-tools.md +95 -95
  32. package/.claude/agents/flow-nexus/workflow.md +83 -83
  33. package/.claude/agents/github/code-review-swarm.md +377 -377
  34. package/.claude/agents/github/github-modes.md +172 -172
  35. package/.claude/agents/github/issue-tracker.md +575 -575
  36. package/.claude/agents/github/multi-repo-swarm.md +552 -552
  37. package/.claude/agents/github/pr-manager.md +437 -437
  38. package/.claude/agents/github/project-board-sync.md +508 -508
  39. package/.claude/agents/github/release-manager.md +604 -604
  40. package/.claude/agents/github/release-swarm.md +582 -582
  41. package/.claude/agents/github/repo-architect.md +397 -397
  42. package/.claude/agents/github/swarm-issue.md +572 -572
  43. package/.claude/agents/github/swarm-pr.md +427 -427
  44. package/.claude/agents/github/sync-coordinator.md +451 -451
  45. package/.claude/agents/github/workflow-automation.md +902 -902
  46. package/.claude/agents/goal/agent.md +815 -815
  47. package/.claude/agents/optimization/benchmark-suite.md +664 -664
  48. package/.claude/agents/optimization/load-balancer.md +430 -430
  49. package/.claude/agents/optimization/performance-monitor.md +671 -671
  50. package/.claude/agents/optimization/resource-allocator.md +673 -673
  51. package/.claude/agents/optimization/topology-optimizer.md +807 -807
  52. package/.claude/agents/payments/agentic-payments.md +126 -126
  53. package/.claude/agents/sona/sona-learning-optimizer.md +74 -74
  54. package/.claude/agents/sparc/architecture.md +698 -698
  55. package/.claude/agents/sparc/pseudocode.md +519 -519
  56. package/.claude/agents/sparc/refinement.md +801 -801
  57. package/.claude/agents/sparc/specification.md +477 -477
  58. package/.claude/agents/specialized/mobile/spec-mobile-react-native.md +224 -224
  59. package/.claude/agents/specialized/spec-mobile-react-native.md +226 -226
  60. package/.claude/agents/sublinear/consensus-coordinator.md +337 -337
  61. package/.claude/agents/sublinear/matrix-optimizer.md +184 -184
  62. package/.claude/agents/sublinear/pagerank-analyzer.md +298 -298
  63. package/.claude/agents/sublinear/performance-optimizer.md +367 -367
  64. package/.claude/agents/sublinear/trading-predictor.md +245 -245
  65. package/.claude/agents/swarm/adaptive-coordinator.md +1126 -1126
  66. package/.claude/agents/swarm/hierarchical-coordinator.md +709 -709
  67. package/.claude/agents/swarm/mesh-coordinator.md +962 -962
  68. package/.claude/agents/templates/automation-smart-agent.md +204 -204
  69. package/.claude/agents/templates/base-template-generator.md +289 -289
  70. package/.claude/agents/templates/coordinator-swarm-init.md +89 -89
  71. package/.claude/agents/templates/github-pr-manager.md +176 -176
  72. package/.claude/agents/templates/implementer-sparc-coder.md +258 -258
  73. package/.claude/agents/templates/memory-coordinator.md +186 -186
  74. package/.claude/agents/templates/orchestrator-task.md +138 -138
  75. package/.claude/agents/templates/performance-analyzer.md +198 -198
  76. package/.claude/agents/templates/sparc-coordinator.md +513 -513
  77. package/.claude/agents/testing/production-validator.md +394 -394
  78. package/.claude/agents/testing/tdd-london-swarm.md +243 -243
  79. package/.claude/agents/v3/aidefence-guardian.md +282 -282
  80. package/.claude/agents/v3/claims-authorizer.md +208 -208
  81. package/.claude/agents/v3/collective-intelligence-coordinator.md +993 -993
  82. package/.claude/agents/v3/ddd-domain-expert.md +220 -220
  83. package/.claude/agents/v3/injection-analyst.md +236 -236
  84. package/.claude/agents/v3/performance-engineer.md +1233 -1233
  85. package/.claude/agents/v3/pii-detector.md +151 -151
  86. package/.claude/agents/v3/reasoningbank-learner.md +213 -213
  87. package/.claude/agents/v3/security-architect-aidefence.md +410 -410
  88. package/.claude/agents/v3/security-architect.md +867 -867
  89. package/.claude/agents/v3/swarm-memory-manager.md +157 -157
  90. package/.claude/agents/v3/v3-integration-architect.md +205 -205
  91. package/.claude/commands/agents/README.md +50 -50
  92. package/.claude/commands/agents/agent-capabilities.md +140 -140
  93. package/.claude/commands/agents/agent-coordination.md +28 -28
  94. package/.claude/commands/agents/agent-spawning.md +28 -28
  95. package/.claude/commands/agents/agent-types.md +216 -216
  96. package/.claude/commands/agents/health.md +139 -139
  97. package/.claude/commands/agents/list.md +100 -100
  98. package/.claude/commands/agents/logs.md +130 -130
  99. package/.claude/commands/agents/metrics.md +122 -122
  100. package/.claude/commands/agents/pool.md +127 -127
  101. package/.claude/commands/agents/spawn.md +140 -140
  102. package/.claude/commands/agents/status.md +115 -115
  103. package/.claude/commands/agents/stop.md +102 -102
  104. package/.claude/commands/analysis/COMMAND_COMPLIANCE_REPORT.md +53 -53
  105. package/.claude/commands/analysis/README.md +9 -9
  106. package/.claude/commands/analysis/bottleneck-detect.md +162 -162
  107. package/.claude/commands/analysis/performance-bottlenecks.md +58 -58
  108. package/.claude/commands/analysis/performance-report.md +25 -25
  109. package/.claude/commands/analysis/token-efficiency.md +44 -44
  110. package/.claude/commands/analysis/token-usage.md +25 -25
  111. package/.claude/commands/automation/README.md +9 -9
  112. package/.claude/commands/automation/auto-agent.md +122 -122
  113. package/.claude/commands/automation/self-healing.md +105 -105
  114. package/.claude/commands/automation/session-memory.md +89 -89
  115. package/.claude/commands/automation/smart-agents.md +72 -72
  116. package/.claude/commands/automation/smart-spawn.md +25 -25
  117. package/.claude/commands/automation/workflow-select.md +25 -25
  118. package/.claude/commands/claude-flow-help.md +103 -103
  119. package/.claude/commands/claude-flow-memory.md +107 -107
  120. package/.claude/commands/claude-flow-swarm.md +205 -205
  121. package/.claude/commands/coordination/README.md +9 -9
  122. package/.claude/commands/coordination/agent-spawn.md +25 -25
  123. package/.claude/commands/coordination/init.md +44 -44
  124. package/.claude/commands/coordination/orchestrate.md +43 -43
  125. package/.claude/commands/coordination/spawn.md +45 -45
  126. package/.claude/commands/coordination/swarm-init.md +85 -85
  127. package/.claude/commands/coordination/task-orchestrate.md +25 -25
  128. package/.claude/commands/github/README.md +11 -11
  129. package/.claude/commands/github/code-review-swarm.md +513 -513
  130. package/.claude/commands/github/code-review.md +25 -25
  131. package/.claude/commands/github/github-modes.md +146 -146
  132. package/.claude/commands/github/github-swarm.md +121 -121
  133. package/.claude/commands/github/issue-tracker.md +291 -291
  134. package/.claude/commands/github/issue-triage.md +25 -25
  135. package/.claude/commands/github/multi-repo-swarm.md +518 -518
  136. package/.claude/commands/github/pr-enhance.md +26 -26
  137. package/.claude/commands/github/pr-manager.md +169 -169
  138. package/.claude/commands/github/project-board-sync.md +470 -470
  139. package/.claude/commands/github/release-manager.md +339 -339
  140. package/.claude/commands/github/release-swarm.md +543 -543
  141. package/.claude/commands/github/repo-analyze.md +25 -25
  142. package/.claude/commands/github/repo-architect.md +366 -366
  143. package/.claude/commands/github/swarm-issue.md +484 -484
  144. package/.claude/commands/github/swarm-pr.md +287 -287
  145. package/.claude/commands/github/sync-coordinator.md +302 -302
  146. package/.claude/commands/github/workflow-automation.md +441 -441
  147. package/.claude/commands/hive-mind/README.md +17 -17
  148. package/.claude/commands/hive-mind/hive-mind-consensus.md +8 -8
  149. package/.claude/commands/hive-mind/hive-mind-init.md +18 -18
  150. package/.claude/commands/hive-mind/hive-mind-memory.md +8 -8
  151. package/.claude/commands/hive-mind/hive-mind-metrics.md +8 -8
  152. package/.claude/commands/hive-mind/hive-mind-resume.md +8 -8
  153. package/.claude/commands/hive-mind/hive-mind-sessions.md +8 -8
  154. package/.claude/commands/hive-mind/hive-mind-spawn.md +21 -21
  155. package/.claude/commands/hive-mind/hive-mind-status.md +8 -8
  156. package/.claude/commands/hive-mind/hive-mind-stop.md +8 -8
  157. package/.claude/commands/hive-mind/hive-mind-wizard.md +8 -8
  158. package/.claude/commands/hive-mind/hive-mind.md +27 -27
  159. package/.claude/commands/hooks/README.md +11 -11
  160. package/.claude/commands/hooks/overview.md +57 -57
  161. package/.claude/commands/hooks/post-edit.md +117 -117
  162. package/.claude/commands/hooks/post-task.md +112 -112
  163. package/.claude/commands/hooks/pre-edit.md +113 -113
  164. package/.claude/commands/hooks/pre-task.md +111 -111
  165. package/.claude/commands/hooks/session-end.md +118 -118
  166. package/.claude/commands/hooks/setup.md +102 -102
  167. package/.claude/commands/memory/README.md +9 -9
  168. package/.claude/commands/memory/memory-persist.md +25 -25
  169. package/.claude/commands/memory/memory-search.md +25 -25
  170. package/.claude/commands/memory/memory-usage.md +25 -25
  171. package/.claude/commands/memory/neural.md +47 -47
  172. package/.claude/commands/monitoring/README.md +9 -9
  173. package/.claude/commands/monitoring/agent-metrics.md +25 -25
  174. package/.claude/commands/monitoring/agents.md +44 -44
  175. package/.claude/commands/monitoring/real-time-view.md +25 -25
  176. package/.claude/commands/monitoring/status.md +46 -46
  177. package/.claude/commands/monitoring/swarm-monitor.md +25 -25
  178. package/.claude/commands/optimization/README.md +9 -9
  179. package/.claude/commands/optimization/auto-topology.md +61 -61
  180. package/.claude/commands/optimization/cache-manage.md +25 -25
  181. package/.claude/commands/optimization/parallel-execute.md +25 -25
  182. package/.claude/commands/optimization/parallel-execution.md +49 -49
  183. package/.claude/commands/optimization/topology-optimize.md +25 -25
  184. package/.claude/commands/pair/README.md +260 -260
  185. package/.claude/commands/pair/commands.md +545 -545
  186. package/.claude/commands/pair/config.md +509 -509
  187. package/.claude/commands/pair/examples.md +511 -511
  188. package/.claude/commands/pair/modes.md +347 -347
  189. package/.claude/commands/pair/session.md +406 -406
  190. package/.claude/commands/pair/start.md +208 -208
  191. package/.claude/commands/sparc/analyzer.md +51 -51
  192. package/.claude/commands/sparc/architect.md +53 -53
  193. package/.claude/commands/sparc/ask.md +97 -97
  194. package/.claude/commands/sparc/batch-executor.md +54 -54
  195. package/.claude/commands/sparc/code.md +89 -89
  196. package/.claude/commands/sparc/coder.md +54 -54
  197. package/.claude/commands/sparc/debug.md +83 -83
  198. package/.claude/commands/sparc/debugger.md +54 -54
  199. package/.claude/commands/sparc/designer.md +53 -53
  200. package/.claude/commands/sparc/devops.md +109 -109
  201. package/.claude/commands/sparc/docs-writer.md +80 -80
  202. package/.claude/commands/sparc/documenter.md +54 -54
  203. package/.claude/commands/sparc/innovator.md +54 -54
  204. package/.claude/commands/sparc/integration.md +83 -83
  205. package/.claude/commands/sparc/mcp.md +117 -117
  206. package/.claude/commands/sparc/memory-manager.md +54 -54
  207. package/.claude/commands/sparc/optimizer.md +54 -54
  208. package/.claude/commands/sparc/orchestrator.md +131 -131
  209. package/.claude/commands/sparc/post-deployment-monitoring-mode.md +83 -83
  210. package/.claude/commands/sparc/refinement-optimization-mode.md +83 -83
  211. package/.claude/commands/sparc/researcher.md +54 -54
  212. package/.claude/commands/sparc/reviewer.md +54 -54
  213. package/.claude/commands/sparc/security-review.md +80 -80
  214. package/.claude/commands/sparc/sparc-modes.md +174 -174
  215. package/.claude/commands/sparc/sparc.md +111 -111
  216. package/.claude/commands/sparc/spec-pseudocode.md +80 -80
  217. package/.claude/commands/sparc/supabase-admin.md +348 -348
  218. package/.claude/commands/sparc/swarm-coordinator.md +54 -54
  219. package/.claude/commands/sparc/tdd.md +54 -54
  220. package/.claude/commands/sparc/tester.md +54 -54
  221. package/.claude/commands/sparc/tutorial.md +79 -79
  222. package/.claude/commands/sparc/workflow-manager.md +54 -54
  223. package/.claude/commands/sparc.md +166 -166
  224. package/.claude/commands/stream-chain/pipeline.md +120 -120
  225. package/.claude/commands/stream-chain/run.md +69 -69
  226. package/.claude/commands/swarm/README.md +15 -15
  227. package/.claude/commands/swarm/analysis.md +95 -95
  228. package/.claude/commands/swarm/development.md +96 -96
  229. package/.claude/commands/swarm/examples.md +168 -168
  230. package/.claude/commands/swarm/maintenance.md +102 -102
  231. package/.claude/commands/swarm/optimization.md +117 -117
  232. package/.claude/commands/swarm/research.md +136 -136
  233. package/.claude/commands/swarm/swarm-analysis.md +8 -8
  234. package/.claude/commands/swarm/swarm-background.md +8 -8
  235. package/.claude/commands/swarm/swarm-init.md +19 -19
  236. package/.claude/commands/swarm/swarm-modes.md +8 -8
  237. package/.claude/commands/swarm/swarm-monitor.md +8 -8
  238. package/.claude/commands/swarm/swarm-spawn.md +19 -19
  239. package/.claude/commands/swarm/swarm-status.md +8 -8
  240. package/.claude/commands/swarm/swarm-strategies.md +8 -8
  241. package/.claude/commands/swarm/swarm.md +87 -87
  242. package/.claude/commands/swarm/testing.md +131 -131
  243. package/.claude/commands/training/README.md +9 -9
  244. package/.claude/commands/training/model-update.md +25 -25
  245. package/.claude/commands/training/neural-patterns.md +107 -107
  246. package/.claude/commands/training/neural-train.md +75 -75
  247. package/.claude/commands/training/pattern-learn.md +25 -25
  248. package/.claude/commands/training/specialization.md +62 -62
  249. package/.claude/commands/truth/start.md +142 -142
  250. package/.claude/commands/verify/check.md +49 -49
  251. package/.claude/commands/verify/start.md +127 -127
  252. package/.claude/commands/workflows/README.md +9 -9
  253. package/.claude/commands/workflows/development.md +77 -77
  254. package/.claude/commands/workflows/research.md +62 -62
  255. package/.claude/commands/workflows/workflow-create.md +25 -25
  256. package/.claude/commands/workflows/workflow-execute.md +25 -25
  257. package/.claude/commands/workflows/workflow-export.md +25 -25
  258. package/.claude/eval/human-relevance-frozen-v1.json +17 -17
  259. package/.claude/evolve-proof/generation-0.json +211 -211
  260. package/.claude/evolve-proof/real-generation-0.json +406 -406
  261. package/.claude/evolve-proof/real-generation-1.json +406 -406
  262. package/.claude/helpers/.helpers-version +1 -1
  263. package/.claude/helpers/README.md +96 -96
  264. package/.claude/helpers/adr-compliance.sh +186 -186
  265. package/.claude/helpers/auto-commit.sh +178 -178
  266. package/.claude/helpers/auto-memory-hook.mjs +430 -430
  267. package/.claude/helpers/checkpoint-manager.sh +251 -251
  268. package/.claude/helpers/daemon-manager.sh +252 -252
  269. package/.claude/helpers/ddd-tracker.sh +144 -144
  270. package/.claude/helpers/github-safe.js +156 -156
  271. package/.claude/helpers/github-setup.sh +45 -45
  272. package/.claude/helpers/guidance-hook.sh +13 -13
  273. package/.claude/helpers/guidance-hooks.sh +102 -102
  274. package/.claude/helpers/health-monitor.sh +108 -108
  275. package/.claude/helpers/helpers.manifest.json +6 -6
  276. package/.claude/helpers/hook-handler.cjs +564 -460
  277. package/.claude/helpers/intelligence.cjs +1058 -1058
  278. package/.claude/helpers/learning-hooks.sh +329 -329
  279. package/.claude/helpers/learning-optimizer.sh +127 -127
  280. package/.claude/helpers/learning-service.mjs +1144 -1144
  281. package/.claude/helpers/memory.js +83 -83
  282. package/.claude/helpers/metrics-db.mjs +503 -503
  283. package/.claude/helpers/pattern-consolidator.sh +86 -86
  284. package/.claude/helpers/perf-worker.sh +160 -160
  285. package/.claude/helpers/post-commit +16 -16
  286. package/.claude/helpers/pre-commit +26 -26
  287. package/.claude/helpers/quick-start.sh +19 -19
  288. package/.claude/helpers/router.js +105 -105
  289. package/.claude/helpers/security-scanner.sh +127 -127
  290. package/.claude/helpers/session.js +157 -157
  291. package/.claude/helpers/setup-mcp.sh +18 -18
  292. package/.claude/helpers/standard-checkpoint-hooks.sh +189 -189
  293. package/.claude/helpers/statusline-hook.sh +21 -21
  294. package/.claude/helpers/statusline.cjs +925 -925
  295. package/.claude/helpers/statusline.js +352 -352
  296. package/.claude/helpers/swarm-comms.sh +353 -353
  297. package/.claude/helpers/swarm-hooks.sh +761 -761
  298. package/.claude/helpers/swarm-monitor.sh +210 -210
  299. package/.claude/helpers/sync-v3-metrics.sh +245 -245
  300. package/.claude/helpers/update-v3-progress.sh +165 -165
  301. package/.claude/helpers/v3-quick-status.sh +57 -57
  302. package/.claude/helpers/v3.sh +110 -110
  303. package/.claude/helpers/validate-v3-config.sh +215 -215
  304. package/.claude/helpers/worker-manager.sh +170 -170
  305. package/.claude/proven-config.json +1 -1
  306. package/.claude/proven-config.manifest.json +37 -37
  307. package/.claude/proven-config.signed.json +41 -41
  308. package/.claude/settings.json +182 -182
  309. package/.claude/skills/agentdb-advanced/SKILL.md +550 -550
  310. package/.claude/skills/agentdb-learning/SKILL.md +545 -545
  311. package/.claude/skills/agentdb-memory-patterns/SKILL.md +339 -339
  312. package/.claude/skills/agentdb-optimization/SKILL.md +509 -509
  313. package/.claude/skills/agentdb-vector-search/SKILL.md +339 -339
  314. package/.claude/skills/browser/SKILL.md +204 -204
  315. package/.claude/skills/dual-mode/README.md +71 -71
  316. package/.claude/skills/dual-mode/dual-collect.md +103 -103
  317. package/.claude/skills/dual-mode/dual-coordinate.md +85 -85
  318. package/.claude/skills/dual-mode/dual-spawn.md +81 -81
  319. package/.claude/skills/flow-nexus-neural/SKILL.md +727 -727
  320. package/.claude/skills/flow-nexus-platform/SKILL.md +1154 -1154
  321. package/.claude/skills/flow-nexus-swarm/SKILL.md +604 -604
  322. package/.claude/skills/github-code-review/SKILL.md +1125 -1125
  323. package/.claude/skills/github-multi-repo/SKILL.md +862 -862
  324. package/.claude/skills/github-project-management/SKILL.md +1262 -1262
  325. package/.claude/skills/github-release-management/SKILL.md +1064 -1064
  326. package/.claude/skills/github-workflow-automation/SKILL.md +1047 -1047
  327. package/.claude/skills/hooks-automation/SKILL.md +1201 -1201
  328. package/.claude/skills/pair-programming/SKILL.md +1202 -1202
  329. package/.claude/skills/reasoningbank-agentdb/SKILL.md +446 -446
  330. package/.claude/skills/reasoningbank-intelligence/SKILL.md +201 -201
  331. package/.claude/skills/skill-builder/SKILL.md +910 -910
  332. package/.claude/skills/sparc-methodology/SKILL.md +1106 -1106
  333. package/.claude/skills/stream-chain/SKILL.md +560 -560
  334. package/.claude/skills/swarm-advanced/SKILL.md +970 -970
  335. package/.claude/skills/swarm-orchestration/SKILL.md +179 -179
  336. package/.claude/skills/v3-cli-modernization/SKILL.md +871 -871
  337. package/.claude/skills/v3-core-implementation/SKILL.md +796 -796
  338. package/.claude/skills/v3-ddd-architecture/SKILL.md +441 -441
  339. package/.claude/skills/v3-integration-deep/SKILL.md +240 -240
  340. package/.claude/skills/v3-mcp-optimization/SKILL.md +776 -776
  341. package/.claude/skills/v3-memory-unification/SKILL.md +173 -173
  342. package/.claude/skills/v3-performance-optimization/SKILL.md +389 -389
  343. package/.claude/skills/v3-security-overhaul/SKILL.md +81 -81
  344. package/.claude/skills/v3-swarm-coordination/SKILL.md +339 -339
  345. package/.claude/skills/verification-quality/SKILL.md +691 -691
  346. package/README.md +419 -419
  347. package/bin/cli.js +314 -314
  348. package/bin/mcp-server.js +224 -224
  349. package/bin/preinstall.cjs +2 -2
  350. package/catalog-manifest.json +2 -2
  351. package/dist/src/benchmarks/gaia-critic.js +24 -24
  352. package/dist/src/business-pods/bbs-budget-tracker.js +53 -53
  353. package/dist/src/commands/announcements.d.ts +17 -0
  354. package/dist/src/commands/announcements.js +0 -0
  355. package/dist/src/commands/completions.js +409 -409
  356. package/dist/src/commands/daemon.js +44 -44
  357. package/dist/src/commands/embeddings.js +26 -26
  358. package/dist/src/commands/funnel.d.ts +10 -5
  359. package/dist/src/commands/funnel.js +206 -7
  360. package/dist/src/commands/hive-mind.js +97 -97
  361. package/dist/src/commands/hooks.js +9 -9
  362. package/dist/src/commands/index.js +4 -0
  363. package/dist/src/commands/init.js +94 -38
  364. package/dist/src/commands/memory.js +85 -1
  365. package/dist/src/commands/ruvector/backup.js +23 -23
  366. package/dist/src/commands/ruvector/benchmark.js +31 -31
  367. package/dist/src/commands/ruvector/import.js +14 -14
  368. package/dist/src/commands/ruvector/init.js +115 -115
  369. package/dist/src/commands/ruvector/migrate.js +99 -99
  370. package/dist/src/commands/ruvector/optimize.js +51 -51
  371. package/dist/src/commands/ruvector/setup.js +624 -624
  372. package/dist/src/commands/ruvector/status.js +38 -38
  373. package/dist/src/commands/spinner.d.ts +16 -0
  374. package/dist/src/commands/spinner.js +329 -0
  375. package/dist/src/config/proven-config.js +2 -2
  376. package/dist/src/funnel/consent.js +3 -0
  377. package/dist/src/funnel/disclosure.d.ts +1 -0
  378. package/dist/src/funnel/disclosure.js +12 -0
  379. package/dist/src/funnel/index.d.ts +2 -1
  380. package/dist/src/funnel/index.js +2 -1
  381. package/dist/src/funnel/payout.d.ts +40 -0
  382. package/dist/src/funnel/payout.js +60 -0
  383. package/dist/src/funnel/types.d.ts +13 -1
  384. package/dist/src/init/claudemd-generator.js +231 -231
  385. package/dist/src/init/executor.js +453 -453
  386. package/dist/src/init/helper-refresh.js +17 -0
  387. package/dist/src/init/helper-signing.d.ts +8 -1
  388. package/dist/src/init/helper-signing.js +9 -2
  389. package/dist/src/init/helpers-generator.js +751 -751
  390. package/dist/src/init/statusline-generator.js +955 -949
  391. package/dist/src/mcp-tools/agentdb-tools.js +15 -15
  392. package/dist/src/mcp-tools/browser-intent-tools.js +19 -19
  393. package/dist/src/memory/graph-edge-writer.js +22 -22
  394. package/dist/src/memory/memory-bridge.d.ts +10 -0
  395. package/dist/src/memory/memory-bridge.js +139 -88
  396. package/dist/src/memory/memory-initializer.d.ts +18 -0
  397. package/dist/src/memory/memory-initializer.js +539 -407
  398. package/dist/src/memory/rabitq-index.js +5 -5
  399. package/dist/src/runtime/headless.js +28 -28
  400. package/dist/src/ruvector/flash-attention.d.ts +195 -0
  401. package/dist/src/ruvector/flash-attention.js +643 -0
  402. package/dist/src/ruvector/moe-router.d.ts +206 -0
  403. package/dist/src/ruvector/moe-router.js +626 -0
  404. package/dist/src/services/distill-tuning.js +7 -7
  405. package/dist/src/services/event-stream.d.ts +25 -0
  406. package/dist/src/services/event-stream.js +27 -0
  407. package/dist/src/services/headless-worker-executor.js +84 -84
  408. package/dist/src/services/loop-worker-runner.d.ts +16 -0
  409. package/dist/src/services/loop-worker-runner.js +34 -0
  410. package/dist/src/services/memory-distillation.js +4 -4
  411. package/dist/src/services/runtime-capabilities.d.ts +22 -0
  412. package/dist/src/services/runtime-capabilities.js +45 -0
  413. package/dist/src/transfer/deploy-seraphine.js +23 -23
  414. package/package.json +134 -133
  415. package/plugins/ruflo-metaharness/.claude-plugin/plugin.json +32 -32
  416. package/plugins/ruflo-metaharness/README.md +72 -72
  417. package/plugins/ruflo-metaharness/agents/metaharness-architect.md +58 -58
  418. package/plugins/ruflo-metaharness/commands/ruflo-metaharness.md +48 -48
  419. package/plugins/ruflo-metaharness/scripts/_darwin.mjs +210 -210
  420. package/plugins/ruflo-metaharness/scripts/_harness.mjs +330 -330
  421. package/plugins/ruflo-metaharness/scripts/_invoke.mjs +231 -231
  422. package/plugins/ruflo-metaharness/scripts/_redblue.mjs +143 -143
  423. package/plugins/ruflo-metaharness/scripts/_similarity.mjs +161 -161
  424. package/plugins/ruflo-metaharness/scripts/_spike-similarity.mjs +223 -223
  425. package/plugins/ruflo-metaharness/scripts/audit-list.mjs +158 -158
  426. package/plugins/ruflo-metaharness/scripts/audit-trend.mjs +272 -272
  427. package/plugins/ruflo-metaharness/scripts/bench-parse-mcp-scan.mjs +146 -146
  428. package/plugins/ruflo-metaharness/scripts/bench-recordpair-overhead.mjs +186 -186
  429. package/plugins/ruflo-metaharness/scripts/bench-similarity.mjs +177 -177
  430. package/plugins/ruflo-metaharness/scripts/bench.mjs +95 -95
  431. package/plugins/ruflo-metaharness/scripts/drift-from-history.mjs +363 -363
  432. package/plugins/ruflo-metaharness/scripts/evolve.mjs +404 -404
  433. package/plugins/ruflo-metaharness/scripts/genome.mjs +80 -80
  434. package/plugins/ruflo-metaharness/scripts/gepa.mjs +153 -153
  435. package/plugins/ruflo-metaharness/scripts/learn.mjs +127 -127
  436. package/plugins/ruflo-metaharness/scripts/mcp-scan.mjs +111 -111
  437. package/plugins/ruflo-metaharness/scripts/mint.mjs +126 -126
  438. package/plugins/ruflo-metaharness/scripts/oia-audit.mjs +228 -228
  439. package/plugins/ruflo-metaharness/scripts/redblue.mjs +286 -286
  440. package/plugins/ruflo-metaharness/scripts/router-parallel-analyze.mjs +250 -250
  441. package/plugins/ruflo-metaharness/scripts/score.mjs +92 -92
  442. package/plugins/ruflo-metaharness/scripts/security-bench.mjs +174 -174
  443. package/plugins/ruflo-metaharness/scripts/similarity.mjs +158 -158
  444. package/plugins/ruflo-metaharness/scripts/smoke.sh +2353 -2353
  445. package/plugins/ruflo-metaharness/scripts/test-graceful-degradation.mjs +165 -165
  446. package/plugins/ruflo-metaharness/scripts/test-mcp-tools.mjs +472 -472
  447. package/plugins/ruflo-metaharness/scripts/test-parallel-pipeline.mjs +204 -204
  448. package/plugins/ruflo-metaharness/scripts/test-pipeline-roundtrip.mjs +586 -586
  449. package/plugins/ruflo-metaharness/scripts/test-similarity.mjs +334 -334
  450. package/plugins/ruflo-metaharness/scripts/test-with-openrouter.mjs +229 -229
  451. package/plugins/ruflo-metaharness/scripts/threat-model.mjs +59 -59
  452. package/plugins/ruflo-metaharness/skills/harness-bench/SKILL.md +64 -64
  453. package/plugins/ruflo-metaharness/skills/harness-drift-from-history/SKILL.md +65 -65
  454. package/plugins/ruflo-metaharness/skills/harness-evolve/SKILL.md +131 -131
  455. package/plugins/ruflo-metaharness/skills/harness-genome/SKILL.md +54 -54
  456. package/plugins/ruflo-metaharness/skills/harness-gepa/SKILL.md +65 -65
  457. package/plugins/ruflo-metaharness/skills/harness-learn/SKILL.md +65 -65
  458. package/plugins/ruflo-metaharness/skills/harness-mcp-scan/SKILL.md +49 -49
  459. package/plugins/ruflo-metaharness/skills/harness-mint/SKILL.md +72 -72
  460. package/plugins/ruflo-metaharness/skills/harness-oia-audit/SKILL.md +79 -79
  461. package/plugins/ruflo-metaharness/skills/harness-score/SKILL.md +66 -66
  462. package/plugins/ruflo-metaharness/skills/harness-security-bench/SKILL.md +101 -101
  463. package/plugins/ruflo-metaharness/skills/harness-similarity/SKILL.md +67 -67
  464. package/plugins/ruflo-metaharness/skills/harness-threat-model/SKILL.md +41 -41
  465. package/scripts/postinstall.cjs +153 -153
@@ -1,64 +1,64 @@
1
- ---
2
- name: harness-bench
3
- description: Manage `@metaharness/darwin` bench suites — `bench create <repo>` scaffolds a JSON suite from a repo's test corpus; `bench verify <suite.json>` checks suite well-formedness. Bench suites are the fixed evaluation corpora that `harness-evolve --bench <suite.json>` scores variants against, decoupling evolution from the repo's natural tests. Degrades gracefully when @metaharness/darwin is absent.
4
- argument-hint: "--op create --repo <path> [--out <path>] | --op verify --suite <path>"
5
- allowed-tools: Bash
6
- ---
7
-
8
- Surfaces `metaharness-darwin bench <create|verify>` — the supporting verb
9
- for `harness-evolve --bench`. Use when you want evolution scored against a
10
- fixed corpus (independent of `npm test`) so champion fitness is comparable
11
- across commits or across forks of the same harness.
12
-
13
- ## When to use
14
-
15
- - Setting up a new evolution pipeline for a repo whose `npm test` is
16
- flaky, slow, or undersized — scaffold a deterministic bench suite once,
17
- then evolve against it repeatedly.
18
- - CI: `bench verify` the checked-in suite on every PR that touches it
19
- (cheap; ~5s).
20
- - Forking a harness to a new domain: copy and edit the suite to retarget
21
- the evaluation without losing comparability to the parent.
22
-
23
- ## Algorithm
24
-
25
- Implementation: [`scripts/bench.mjs`](../../scripts/bench.mjs).
26
-
27
- ### `--op create`
28
- 1. Resolve `--repo` path; reject if missing.
29
- 2. Shell to `metaharness-darwin bench create <repo> [--out <suite.json>]`.
30
- 3. Default output path: `<repo>/.metaharness/bench/suite.json` (chosen by upstream).
31
- 4. Suite shape (per upstream): array of `{ input, expectedOutput, weight }` tasks
32
- derived from existing test cases.
33
-
34
- ### `--op verify`
35
- 1. Resolve `--suite` path; reject if missing.
36
- 2. Shell to `metaharness-darwin bench verify <suite.json>`.
37
- 3. Exit 1 if any task malformed (upstream's signal).
38
-
39
- ## Output shape
40
-
41
- ```json
42
- {
43
- "success": true,
44
- "data": {
45
- "op": "verify",
46
- "taskCount": 42,
47
- "wellFormed": true,
48
- "durationMs": 870
49
- }
50
- }
51
- ```
52
-
53
- ## Exit codes
54
-
55
- | Code | Meaning |
56
- |---|---|
57
- | 0 | OK (or degraded — Darwin absent) |
58
- | 1 | `--op verify` and suite malformed |
59
- | 2 | Config error or upstream invocation failure |
60
-
61
- ## Graceful degradation
62
-
63
- When `@metaharness/darwin` is absent, emits the standard `{degraded: true,
64
- reason: 'metaharness-darwin-not-available'}` payload and exits 0.
1
+ ---
2
+ name: harness-bench
3
+ description: Manage `@metaharness/darwin` bench suites — `bench create <repo>` scaffolds a JSON suite from a repo's test corpus; `bench verify <suite.json>` checks suite well-formedness. Bench suites are the fixed evaluation corpora that `harness-evolve --bench <suite.json>` scores variants against, decoupling evolution from the repo's natural tests. Degrades gracefully when @metaharness/darwin is absent.
4
+ argument-hint: "--op create --repo <path> [--out <path>] | --op verify --suite <path>"
5
+ allowed-tools: Bash
6
+ ---
7
+
8
+ Surfaces `metaharness-darwin bench <create|verify>` — the supporting verb
9
+ for `harness-evolve --bench`. Use when you want evolution scored against a
10
+ fixed corpus (independent of `npm test`) so champion fitness is comparable
11
+ across commits or across forks of the same harness.
12
+
13
+ ## When to use
14
+
15
+ - Setting up a new evolution pipeline for a repo whose `npm test` is
16
+ flaky, slow, or undersized — scaffold a deterministic bench suite once,
17
+ then evolve against it repeatedly.
18
+ - CI: `bench verify` the checked-in suite on every PR that touches it
19
+ (cheap; ~5s).
20
+ - Forking a harness to a new domain: copy and edit the suite to retarget
21
+ the evaluation without losing comparability to the parent.
22
+
23
+ ## Algorithm
24
+
25
+ Implementation: [`scripts/bench.mjs`](../../scripts/bench.mjs).
26
+
27
+ ### `--op create`
28
+ 1. Resolve `--repo` path; reject if missing.
29
+ 2. Shell to `metaharness-darwin bench create <repo> [--out <suite.json>]`.
30
+ 3. Default output path: `<repo>/.metaharness/bench/suite.json` (chosen by upstream).
31
+ 4. Suite shape (per upstream): array of `{ input, expectedOutput, weight }` tasks
32
+ derived from existing test cases.
33
+
34
+ ### `--op verify`
35
+ 1. Resolve `--suite` path; reject if missing.
36
+ 2. Shell to `metaharness-darwin bench verify <suite.json>`.
37
+ 3. Exit 1 if any task malformed (upstream's signal).
38
+
39
+ ## Output shape
40
+
41
+ ```json
42
+ {
43
+ "success": true,
44
+ "data": {
45
+ "op": "verify",
46
+ "taskCount": 42,
47
+ "wellFormed": true,
48
+ "durationMs": 870
49
+ }
50
+ }
51
+ ```
52
+
53
+ ## Exit codes
54
+
55
+ | Code | Meaning |
56
+ |---|---|
57
+ | 0 | OK (or degraded — Darwin absent) |
58
+ | 1 | `--op verify` and suite malformed |
59
+ | 2 | Config error or upstream invocation failure |
60
+
61
+ ## Graceful degradation
62
+
63
+ When `@metaharness/darwin` is absent, emits the standard `{degraded: true,
64
+ reason: 'metaharness-darwin-not-available'}` payload and exits 0.
@@ -1,65 +1,65 @@
1
- ---
2
- name: harness-drift-from-history
3
- description: One-command drift detection. Composes audit-list + oia-audit + audit-trend into a single primitive — finds the most recent audit in `metaharness-audit` namespace, runs a fresh audit against the current repo, diffs them via ADR-152 §3.1 similarity, and alerts when structural distance crosses `--threshold`. Iter 53 of ADR-150 deep integration.
4
- argument-hint: "[--path .] [--baseline-since 7d] [--threshold 0.95] [--dry-run] [--format json|table]"
5
- allowed-tools: Bash
6
- ---
7
-
8
- The natural ops question after running `oia-audit` weekly is "did anything drift?" Before this skill, the answer required a three-step sequence:
9
-
10
- ```bash
11
- npx ruflo metaharness audit-list --format json # → pick a key by hand
12
- npx ruflo metaharness oia-audit --format json > /tmp/curr.json
13
- npx ruflo metaharness audit-trend \
14
- --baseline-key <picked-key> --current /tmp/curr.json \
15
- --alert-on-distance-below 0.95
16
- ```
17
-
18
- This skill collapses it into one command:
19
-
20
- ```bash
21
- npx ruflo metaharness drift-from-history --threshold 0.95
22
- ```
23
-
24
- ## What it does
25
-
26
- 1. Lists records from `metaharness-audit` namespace via audit-list.mjs
27
- 2. Picks the most recent record by `startedAt` (or `--baseline-since 7d` skips anything newer than 7 days)
28
- 3. Runs a fresh `oia-audit` against the current path
29
- 4. Diffs the two via audit-trend, applying `--alert-on-distance-below ${threshold}`
30
- 5. Returns the structured drift report
31
-
32
- ## Architectural constraint inheritance (ADR-150)
33
-
34
- | Constraint | How drift-from-history satisfies it |
35
- |---|---|
36
- | Removable | Pure subprocess composition over existing scripts — no new `@metaharness/*` import |
37
- | Optional | If oia-audit reports `degraded:true`, this skill exits 3 with a degraded payload |
38
- | Graceful | Empty audit history → exit 2 with hint to seed it; never crashes |
39
- | CI-gate | Smoke step 17z16 anchors the dispatcher entry + subcommand listing |
40
-
41
- ## Exit codes
42
-
43
- - 0 — similarity ≥ threshold (or threshold not crossed)
44
- - 1 — drift detected: similarity < threshold (alert fired)
45
- - 2 — config error (no history, audit-list failed)
46
- - 3 — upstream metaharness absent (degraded payload returned)
47
-
48
- ## Example
49
-
50
- ```bash
51
- $ npx ruflo metaharness drift-from-history --threshold 0.95
52
- # drift-from-history
53
-
54
- Baseline: audit-2026-06-16T22-58-47-840Z
55
- Current: 2026-06-16T23:05:02.231Z
56
-
57
- Structural similarity: 1 (near-identical)
58
- Distance: 0
59
-
60
- ✓ similarity ≥ 0.95 — OK
61
- ```
62
-
63
- ## Implementation
64
-
65
- [`scripts/drift-from-history.mjs`](../../scripts/drift-from-history.mjs)
1
+ ---
2
+ name: harness-drift-from-history
3
+ description: One-command drift detection. Composes audit-list + oia-audit + audit-trend into a single primitive — finds the most recent audit in `metaharness-audit` namespace, runs a fresh audit against the current repo, diffs them via ADR-152 §3.1 similarity, and alerts when structural distance crosses `--threshold`. Iter 53 of ADR-150 deep integration.
4
+ argument-hint: "[--path .] [--baseline-since 7d] [--threshold 0.95] [--dry-run] [--format json|table]"
5
+ allowed-tools: Bash
6
+ ---
7
+
8
+ The natural ops question after running `oia-audit` weekly is "did anything drift?" Before this skill, the answer required a three-step sequence:
9
+
10
+ ```bash
11
+ npx ruflo metaharness audit-list --format json # → pick a key by hand
12
+ npx ruflo metaharness oia-audit --format json > /tmp/curr.json
13
+ npx ruflo metaharness audit-trend \
14
+ --baseline-key <picked-key> --current /tmp/curr.json \
15
+ --alert-on-distance-below 0.95
16
+ ```
17
+
18
+ This skill collapses it into one command:
19
+
20
+ ```bash
21
+ npx ruflo metaharness drift-from-history --threshold 0.95
22
+ ```
23
+
24
+ ## What it does
25
+
26
+ 1. Lists records from `metaharness-audit` namespace via audit-list.mjs
27
+ 2. Picks the most recent record by `startedAt` (or `--baseline-since 7d` skips anything newer than 7 days)
28
+ 3. Runs a fresh `oia-audit` against the current path
29
+ 4. Diffs the two via audit-trend, applying `--alert-on-distance-below ${threshold}`
30
+ 5. Returns the structured drift report
31
+
32
+ ## Architectural constraint inheritance (ADR-150)
33
+
34
+ | Constraint | How drift-from-history satisfies it |
35
+ |---|---|
36
+ | Removable | Pure subprocess composition over existing scripts — no new `@metaharness/*` import |
37
+ | Optional | If oia-audit reports `degraded:true`, this skill exits 3 with a degraded payload |
38
+ | Graceful | Empty audit history → exit 2 with hint to seed it; never crashes |
39
+ | CI-gate | Smoke step 17z16 anchors the dispatcher entry + subcommand listing |
40
+
41
+ ## Exit codes
42
+
43
+ - 0 — similarity ≥ threshold (or threshold not crossed)
44
+ - 1 — drift detected: similarity < threshold (alert fired)
45
+ - 2 — config error (no history, audit-list failed)
46
+ - 3 — upstream metaharness absent (degraded payload returned)
47
+
48
+ ## Example
49
+
50
+ ```bash
51
+ $ npx ruflo metaharness drift-from-history --threshold 0.95
52
+ # drift-from-history
53
+
54
+ Baseline: audit-2026-06-16T22-58-47-840Z
55
+ Current: 2026-06-16T23:05:02.231Z
56
+
57
+ Structural similarity: 1 (near-identical)
58
+ Distance: 0
59
+
60
+ ✓ similarity ≥ 0.95 — OK
61
+ ```
62
+
63
+ ## Implementation
64
+
65
+ [`scripts/drift-from-history.mjs`](../../scripts/drift-from-history.mjs)
@@ -1,131 +1,131 @@
1
- ---
2
- name: harness-evolve
3
- description: Run `@metaharness/darwin evolve <repo>` to mutate a harness's seven policy surfaces (planner/contextBuilder/reviewer/retryPolicy/toolPolicy/memoryPolicy/scorePolicy), sandbox-score each variant, and promote only measured wins. The model is frozen; the harness evolves. Closes the loop ADR-150 opens (score+genome describe; evolve changes). Degrades gracefully when @metaharness/darwin is absent (ADR-150 + ADR-153 architectural constraints).
4
- argument-hint: "--repo <path> [--generations 3] [--children 3] [--concurrency 2] [--sandbox real|mock|agent] [--selection pareto|quality-diversity|...] [--mutator deterministic|ruvllm] [--diagnose] [--confirm]"
5
- allowed-tools: Bash
6
- ---
7
-
8
- Surfaces the upstream `metaharness-darwin evolve` CLI as a ruflo skill. The
9
- **write** layer that pairs with ADR-150's read layer (score / genome /
10
- mcp-scan / threat-model / oia-audit). Use when you have a harness whose
11
- readiness scores are flat and you want to discover *which* surface mutation
12
- moves them — without retraining the foundation model.
13
-
14
- ## When to use
15
-
16
- - A `harness-score` result is below target and you don't know which policy
17
- surface is responsible.
18
- - You're seeding a harness for a new vertical and want to find a good
19
- starting configuration empirically rather than hand-tuning.
20
- - You're comparing your hand-tuned harness against an evolved baseline
21
- (treat darwin's champion as the strawman).
22
-
23
- ## When NOT to use
24
-
25
- - For continuous background optimization. Darwin Mode is human-initiated.
26
- Wire it into CI for one-shot exploration, not for autonomous self-modification.
27
- - For ruflo itself in CI. ADR-153 §5 explicitly rejects auto-evolving ruflo
28
- — the CI gate verifies graceful degradation, not convergence.
29
-
30
- ## Algorithm
31
-
32
- Implementation: [`scripts/evolve.mjs`](../../scripts/evolve.mjs).
33
-
34
- 1. Validate args (`--repo` exists, caps on `--generations` ≤ 50, `--children`
35
- ≤ 20, `--concurrency` ≤ 8, sandbox/selection/mutator are known values).
36
- 2. Without `--confirm`: print plan + exit 0 (mirrors `harness-mint` safety
37
- convention; defense in depth over the upstream `safety.ts` checks).
38
- 3. With `--confirm`: shell to `npx -y @metaharness/darwin@~0.8.0 metaharness-darwin evolve <repo> ...`
39
- via the shared `_darwin.mjs` async helper. Per-generation progress is
40
- forwarded to stderr; final champion JSON is captured from stdout.
41
- 4. Compute timeout from `generations × children × per-variant` (per-variant
42
- ≈ 60s real, ≈ 2s mock). Caller may override with `--timeout-ms`.
43
- 5. Honor upstream exit code 99 — propagate as "safety-disqualified", do not
44
- remap. This is a designed-in tripwire (a variant tripped `inspectVariant`
45
- for secrets / shell-out / network / dynamic-eval). See ADR-153 §"Safety model".
46
- 6. Optional `--alert-on-no-improvement`: exit 1 when champion ≤ parent.
47
-
48
- ## The seven mutation surfaces
49
-
50
- | Surface | What it owns |
51
- |---|---|
52
- | `planner` | task decomposition / step ordering |
53
- | `contextBuilder` | what gets fed into the prompt |
54
- | `reviewer` | self-critique / output verification |
55
- | `retryPolicy` | when + how to retry on failure |
56
- | `toolPolicy` | which tools the agent may use, under which conditions |
57
- | `memoryPolicy` | what to persist, recall, forget |
58
- | `scorePolicy` | how the agent grades its own output |
59
-
60
- One mutation per variant. Multi-surface mutations are not allowed (causal
61
- attribution stays clean).
62
-
63
- ## Output
64
-
65
- Reports land under `<repo>/.metaharness/`:
66
-
67
- ```
68
- .metaharness/
69
- archive.json # full lineage tree (sampling next gen draws from this)
70
- lineage.json # parent→child edges only
71
- variants/<id>/ # per-variant code (kept for audit)
72
- runs/<id>/ # per-variant sandbox test output
73
- reports/winner.json # final champion + score delta vs parent
74
- ```
75
-
76
- Skill stdout = JSON `{success, data: {champion, plan, durationMs, improved}}`
77
- (plus `data.diagnosis` when `--diagnose` is passed — see below).
78
-
79
- ## Failure diagnosis (`--diagnose`)
80
-
81
- GEPA's key trick is natural-language failure diagnosis from execution traces
82
- feeding the next mutation — not just scalar fitness. `--diagnose` adds a
83
- modest slice of that: after the evolution completes, the losing / failed
84
- variants' transcripts are run through darwin's GEPA library ops
85
- (`analyzeTranscript` + `classifyFailure`, via the shared `importGepa`
86
- resolver in `scripts/_darwin.mjs`) and a `diagnosis` section is appended to
87
- the emitted JSON:
88
-
89
- ```json
90
- "diagnosis": {
91
- "available": true,
92
- "scope": "losing-variants",
93
- "variants": [
94
- { "id": "g1_v0", "transcripts": 2,
95
- "failureClasses": { "exploration-loop": 1, "edit-mechanics": 1 },
96
- "dominantClass": "exploration-loop" }
97
- ],
98
- "totals": { "exploration-loop": 1, "edit-mechanics": 1 }
99
- }
100
- ```
101
-
102
- Upstream shape caveats (verified against `@metaharness/darwin@0.8.0`):
103
-
104
- - `metaharness-darwin evolve --json` prints a TEXT leaderboard — the stdout
105
- carries no JSON and no transcripts. Per-variant run records live at
106
- `<repo>/.metaharness/runs/<id>.json`.
107
- - Those run records hold sandbox exec traces (`{taskId, exitCode, stdout,
108
- stderr}`), which are NOT GEPA `{actionRaw, obs}` transcripts. Diagnosis
109
- therefore uses GEPA-shaped transcripts when a run record embeds them
110
- (agent sandbox / future upstream), falls back to the champion's transcript,
111
- and otherwise emits `diagnosis: {available: false, reason, traceSummary}`
112
- where `traceSummary` is a mechanical per-variant tally (tasks / failed /
113
- timedOut / blockedActions).
114
- - `--diagnose` NEVER fails the run — any internal error degrades to
115
- `{available: false, reason: "diagnosis-failed: ..."}`.
116
-
117
- ## Exit codes
118
-
119
- | Code | Meaning |
120
- |---|---|
121
- | 0 | Evolved OK, or dry-run, or degraded (Darwin absent) |
122
- | 1 | `--alert-on-no-improvement` and champion did not beat parent |
123
- | 2 | Config error or evolution infrastructure failure |
124
- | 99 | Upstream "safety-disqualified" (PROPAGATED, not remapped) |
125
-
126
- ## Graceful degradation (ADR-150 constraint 3 + ADR-153)
127
-
128
- When `@metaharness/darwin` is not installed, the script emits
129
- `{degraded: true, reason: 'metaharness-darwin-not-available', hint: ...}`
130
- and exits 0. ruflo continues to function. CI's
131
- `no-metaharness-smoke.yml`-style job asserts this path.
1
+ ---
2
+ name: harness-evolve
3
+ description: Run `@metaharness/darwin evolve <repo>` to mutate a harness's seven policy surfaces (planner/contextBuilder/reviewer/retryPolicy/toolPolicy/memoryPolicy/scorePolicy), sandbox-score each variant, and promote only measured wins. The model is frozen; the harness evolves. Closes the loop ADR-150 opens (score+genome describe; evolve changes). Degrades gracefully when @metaharness/darwin is absent (ADR-150 + ADR-153 architectural constraints).
4
+ argument-hint: "--repo <path> [--generations 3] [--children 3] [--concurrency 2] [--sandbox real|mock|agent] [--selection pareto|quality-diversity|...] [--mutator deterministic|ruvllm] [--diagnose] [--confirm]"
5
+ allowed-tools: Bash
6
+ ---
7
+
8
+ Surfaces the upstream `metaharness-darwin evolve` CLI as a ruflo skill. The
9
+ **write** layer that pairs with ADR-150's read layer (score / genome /
10
+ mcp-scan / threat-model / oia-audit). Use when you have a harness whose
11
+ readiness scores are flat and you want to discover *which* surface mutation
12
+ moves them — without retraining the foundation model.
13
+
14
+ ## When to use
15
+
16
+ - A `harness-score` result is below target and you don't know which policy
17
+ surface is responsible.
18
+ - You're seeding a harness for a new vertical and want to find a good
19
+ starting configuration empirically rather than hand-tuning.
20
+ - You're comparing your hand-tuned harness against an evolved baseline
21
+ (treat darwin's champion as the strawman).
22
+
23
+ ## When NOT to use
24
+
25
+ - For continuous background optimization. Darwin Mode is human-initiated.
26
+ Wire it into CI for one-shot exploration, not for autonomous self-modification.
27
+ - For ruflo itself in CI. ADR-153 §5 explicitly rejects auto-evolving ruflo
28
+ — the CI gate verifies graceful degradation, not convergence.
29
+
30
+ ## Algorithm
31
+
32
+ Implementation: [`scripts/evolve.mjs`](../../scripts/evolve.mjs).
33
+
34
+ 1. Validate args (`--repo` exists, caps on `--generations` ≤ 50, `--children`
35
+ ≤ 20, `--concurrency` ≤ 8, sandbox/selection/mutator are known values).
36
+ 2. Without `--confirm`: print plan + exit 0 (mirrors `harness-mint` safety
37
+ convention; defense in depth over the upstream `safety.ts` checks).
38
+ 3. With `--confirm`: shell to `npx -y @metaharness/darwin@~0.8.0 metaharness-darwin evolve <repo> ...`
39
+ via the shared `_darwin.mjs` async helper. Per-generation progress is
40
+ forwarded to stderr; final champion JSON is captured from stdout.
41
+ 4. Compute timeout from `generations × children × per-variant` (per-variant
42
+ ≈ 60s real, ≈ 2s mock). Caller may override with `--timeout-ms`.
43
+ 5. Honor upstream exit code 99 — propagate as "safety-disqualified", do not
44
+ remap. This is a designed-in tripwire (a variant tripped `inspectVariant`
45
+ for secrets / shell-out / network / dynamic-eval). See ADR-153 §"Safety model".
46
+ 6. Optional `--alert-on-no-improvement`: exit 1 when champion ≤ parent.
47
+
48
+ ## The seven mutation surfaces
49
+
50
+ | Surface | What it owns |
51
+ |---|---|
52
+ | `planner` | task decomposition / step ordering |
53
+ | `contextBuilder` | what gets fed into the prompt |
54
+ | `reviewer` | self-critique / output verification |
55
+ | `retryPolicy` | when + how to retry on failure |
56
+ | `toolPolicy` | which tools the agent may use, under which conditions |
57
+ | `memoryPolicy` | what to persist, recall, forget |
58
+ | `scorePolicy` | how the agent grades its own output |
59
+
60
+ One mutation per variant. Multi-surface mutations are not allowed (causal
61
+ attribution stays clean).
62
+
63
+ ## Output
64
+
65
+ Reports land under `<repo>/.metaharness/`:
66
+
67
+ ```
68
+ .metaharness/
69
+ archive.json # full lineage tree (sampling next gen draws from this)
70
+ lineage.json # parent→child edges only
71
+ variants/<id>/ # per-variant code (kept for audit)
72
+ runs/<id>/ # per-variant sandbox test output
73
+ reports/winner.json # final champion + score delta vs parent
74
+ ```
75
+
76
+ Skill stdout = JSON `{success, data: {champion, plan, durationMs, improved}}`
77
+ (plus `data.diagnosis` when `--diagnose` is passed — see below).
78
+
79
+ ## Failure diagnosis (`--diagnose`)
80
+
81
+ GEPA's key trick is natural-language failure diagnosis from execution traces
82
+ feeding the next mutation — not just scalar fitness. `--diagnose` adds a
83
+ modest slice of that: after the evolution completes, the losing / failed
84
+ variants' transcripts are run through darwin's GEPA library ops
85
+ (`analyzeTranscript` + `classifyFailure`, via the shared `importGepa`
86
+ resolver in `scripts/_darwin.mjs`) and a `diagnosis` section is appended to
87
+ the emitted JSON:
88
+
89
+ ```json
90
+ "diagnosis": {
91
+ "available": true,
92
+ "scope": "losing-variants",
93
+ "variants": [
94
+ { "id": "g1_v0", "transcripts": 2,
95
+ "failureClasses": { "exploration-loop": 1, "edit-mechanics": 1 },
96
+ "dominantClass": "exploration-loop" }
97
+ ],
98
+ "totals": { "exploration-loop": 1, "edit-mechanics": 1 }
99
+ }
100
+ ```
101
+
102
+ Upstream shape caveats (verified against `@metaharness/darwin@0.8.0`):
103
+
104
+ - `metaharness-darwin evolve --json` prints a TEXT leaderboard — the stdout
105
+ carries no JSON and no transcripts. Per-variant run records live at
106
+ `<repo>/.metaharness/runs/<id>.json`.
107
+ - Those run records hold sandbox exec traces (`{taskId, exitCode, stdout,
108
+ stderr}`), which are NOT GEPA `{actionRaw, obs}` transcripts. Diagnosis
109
+ therefore uses GEPA-shaped transcripts when a run record embeds them
110
+ (agent sandbox / future upstream), falls back to the champion's transcript,
111
+ and otherwise emits `diagnosis: {available: false, reason, traceSummary}`
112
+ where `traceSummary` is a mechanical per-variant tally (tasks / failed /
113
+ timedOut / blockedActions).
114
+ - `--diagnose` NEVER fails the run — any internal error degrades to
115
+ `{available: false, reason: "diagnosis-failed: ..."}`.
116
+
117
+ ## Exit codes
118
+
119
+ | Code | Meaning |
120
+ |---|---|
121
+ | 0 | Evolved OK, or dry-run, or degraded (Darwin absent) |
122
+ | 1 | `--alert-on-no-improvement` and champion did not beat parent |
123
+ | 2 | Config error or evolution infrastructure failure |
124
+ | 99 | Upstream "safety-disqualified" (PROPAGATED, not remapped) |
125
+
126
+ ## Graceful degradation (ADR-150 constraint 3 + ADR-153)
127
+
128
+ When `@metaharness/darwin` is not installed, the script emits
129
+ `{degraded: true, reason: 'metaharness-darwin-not-available', hint: ...}`
130
+ and exits 0. ruflo continues to function. CI's
131
+ `no-metaharness-smoke.yml`-style job asserts this path.
@@ -1,54 +1,54 @@
1
- ---
2
- name: harness-genome
3
- description: 7-section repo readiness report from `metaharness genome <path>`. Returns repo_type / agent_topology / risk_score / mcp_surface / test_confidence / publish_readiness. Pure-read; degrades gracefully (ADR-150).
4
- argument-hint: "[--path .] [--alert-on-risk-above 0.5] [--format table|json]"
5
- allowed-tools: Bash
6
- ---
7
-
8
- Companion to `harness-score`. Where score is a 5-dimension numeric
9
- scorecard, genome is a 7-section categorical/numeric report covering
10
- repo type, agent topology recommendations, risk score (0-1), MCP
11
- surface area, test confidence (0-1), and publish readiness (0-1).
12
-
13
- ## Algorithm
14
-
15
- Implementation: [`scripts/genome.mjs`](../../scripts/genome.mjs).
16
-
17
- 1. Shell out to `npx metaharness genome <path> --json` (60s hard timeout).
18
- 2. Parse the shape: `{ repo_type, agent_topology[], risk_score,
19
- mcp_surface, test_confidence, publish_readiness }`.
20
- 3. If `--alert-on-risk-above N`: exit 1 when `risk_score > N`.
21
- 4. Output JSON (default) or markdown.
22
-
23
- ## Phase-0 baseline (ruflo, measured 2026-06-16)
24
-
25
- ```
26
- {
27
- "repo_type": "node_mcp_ci",
28
- "agent_topology": ["maintainer", "tester", "security", "release"],
29
- "risk_score": 0.27,
30
- "mcp_surface": "remote",
31
- "test_confidence": 0.8,
32
- "publish_readiness": 0.9
33
- }
34
- ```
35
-
36
- Ruflo's `risk_score: 0.27` is low (good). `publish_readiness: 0.9` is
37
- high. The `mcp_surface: "remote"` reflects that ruflo's MCP servers are
38
- hosted, not bundled.
39
-
40
- ## When to use
41
-
42
- - Pre-mint review: "before scaffolding a custom harness from this repo,
43
- should we?" — genome answers it categorically.
44
- - Drift detection: capture genome snapshots over time, diff via
45
- cost-diff-style tooling to spot when `agent_topology` recommendations
46
- drift away from a deliberate architecture choice.
47
- - CI gate: `--alert-on-risk-above 0.5` fails the build when the repo's
48
- risk profile crosses a threshold.
49
-
50
- ## Pairs with
51
-
52
- - `harness-score` — numeric readiness
53
- - `harness-mcp-scan` — static MCP security findings
54
- - `harness-threat-model` — enterprise-review-grade threat model
1
+ ---
2
+ name: harness-genome
3
+ description: 7-section repo readiness report from `metaharness genome <path>`. Returns repo_type / agent_topology / risk_score / mcp_surface / test_confidence / publish_readiness. Pure-read; degrades gracefully (ADR-150).
4
+ argument-hint: "[--path .] [--alert-on-risk-above 0.5] [--format table|json]"
5
+ allowed-tools: Bash
6
+ ---
7
+
8
+ Companion to `harness-score`. Where score is a 5-dimension numeric
9
+ scorecard, genome is a 7-section categorical/numeric report covering
10
+ repo type, agent topology recommendations, risk score (0-1), MCP
11
+ surface area, test confidence (0-1), and publish readiness (0-1).
12
+
13
+ ## Algorithm
14
+
15
+ Implementation: [`scripts/genome.mjs`](../../scripts/genome.mjs).
16
+
17
+ 1. Shell out to `npx metaharness genome <path> --json` (60s hard timeout).
18
+ 2. Parse the shape: `{ repo_type, agent_topology[], risk_score,
19
+ mcp_surface, test_confidence, publish_readiness }`.
20
+ 3. If `--alert-on-risk-above N`: exit 1 when `risk_score > N`.
21
+ 4. Output JSON (default) or markdown.
22
+
23
+ ## Phase-0 baseline (ruflo, measured 2026-06-16)
24
+
25
+ ```
26
+ {
27
+ "repo_type": "node_mcp_ci",
28
+ "agent_topology": ["maintainer", "tester", "security", "release"],
29
+ "risk_score": 0.27,
30
+ "mcp_surface": "remote",
31
+ "test_confidence": 0.8,
32
+ "publish_readiness": 0.9
33
+ }
34
+ ```
35
+
36
+ Ruflo's `risk_score: 0.27` is low (good). `publish_readiness: 0.9` is
37
+ high. The `mcp_surface: "remote"` reflects that ruflo's MCP servers are
38
+ hosted, not bundled.
39
+
40
+ ## When to use
41
+
42
+ - Pre-mint review: "before scaffolding a custom harness from this repo,
43
+ should we?" — genome answers it categorically.
44
+ - Drift detection: capture genome snapshots over time, diff via
45
+ cost-diff-style tooling to spot when `agent_topology` recommendations
46
+ drift away from a deliberate architecture choice.
47
+ - CI gate: `--alert-on-risk-above 0.5` fails the build when the repo's
48
+ risk profile crosses a threshold.
49
+
50
+ ## Pairs with
51
+
52
+ - `harness-score` — numeric readiness
53
+ - `harness-mcp-scan` — static MCP security findings
54
+ - `harness-threat-model` — enterprise-review-grade threat model