@claude-flow/cli 3.32.9 → 3.32.10
This diff represents the content of publicly available package versions that have been released to one of the supported registries. The information contained in this diff is provided for informational purposes only and reflects changes between package versions as they appear in their respective public registries.
- package/.claude/agents/analysis/analyze-code-quality.md +178 -178
- package/.claude/agents/analysis/code-analyzer.md +209 -209
- package/.claude/agents/analysis/code-review/analyze-code-quality.md +178 -178
- package/.claude/agents/architecture/arch-system-design.md +156 -156
- package/.claude/agents/architecture/system-design/arch-system-design.md +154 -154
- package/.claude/agents/browser/browser-agent.yaml +182 -182
- package/.claude/agents/consensus/byzantine-coordinator.md +62 -62
- package/.claude/agents/consensus/crdt-synchronizer.md +996 -996
- package/.claude/agents/consensus/gossip-coordinator.md +62 -62
- package/.claude/agents/consensus/performance-benchmarker.md +850 -850
- package/.claude/agents/consensus/quorum-manager.md +822 -822
- package/.claude/agents/consensus/raft-manager.md +62 -62
- package/.claude/agents/consensus/security-manager.md +621 -621
- package/.claude/agents/core/planner.md +374 -374
- package/.claude/agents/custom/test-long-runner.md +44 -44
- package/.claude/agents/data/data-ml-model.md +444 -444
- package/.claude/agents/data/ml/data-ml-model.md +192 -192
- package/.claude/agents/development/backend/dev-backend-api.md +141 -141
- package/.claude/agents/development/dev-backend-api.md +344 -344
- package/.claude/agents/devops/ci-cd/ops-cicd-github.md +163 -163
- package/.claude/agents/devops/ops-cicd-github.md +164 -164
- package/.claude/agents/documentation/api-docs/docs-api-openapi.md +173 -173
- package/.claude/agents/documentation/docs-api-openapi.md +354 -354
- package/.claude/agents/flow-nexus/app-store.md +87 -87
- package/.claude/agents/flow-nexus/authentication.md +68 -68
- package/.claude/agents/flow-nexus/challenges.md +80 -80
- package/.claude/agents/flow-nexus/neural-network.md +87 -87
- package/.claude/agents/flow-nexus/payments.md +82 -82
- package/.claude/agents/flow-nexus/sandbox.md +75 -75
- package/.claude/agents/flow-nexus/swarm.md +75 -75
- package/.claude/agents/flow-nexus/user-tools.md +95 -95
- package/.claude/agents/flow-nexus/workflow.md +83 -83
- package/.claude/agents/github/code-review-swarm.md +377 -377
- package/.claude/agents/github/github-modes.md +172 -172
- package/.claude/agents/github/issue-tracker.md +575 -575
- package/.claude/agents/github/multi-repo-swarm.md +552 -552
- package/.claude/agents/github/pr-manager.md +437 -437
- package/.claude/agents/github/project-board-sync.md +508 -508
- package/.claude/agents/github/release-manager.md +604 -604
- package/.claude/agents/github/release-swarm.md +582 -582
- package/.claude/agents/github/repo-architect.md +397 -397
- package/.claude/agents/github/swarm-issue.md +572 -572
- package/.claude/agents/github/swarm-pr.md +427 -427
- package/.claude/agents/github/sync-coordinator.md +451 -451
- package/.claude/agents/github/workflow-automation.md +902 -902
- package/.claude/agents/goal/agent.md +815 -815
- package/.claude/agents/optimization/benchmark-suite.md +664 -664
- package/.claude/agents/optimization/load-balancer.md +430 -430
- package/.claude/agents/optimization/performance-monitor.md +671 -671
- package/.claude/agents/optimization/resource-allocator.md +673 -673
- package/.claude/agents/optimization/topology-optimizer.md +807 -807
- package/.claude/agents/payments/agentic-payments.md +126 -126
- package/.claude/agents/sona/sona-learning-optimizer.md +74 -74
- package/.claude/agents/sparc/architecture.md +698 -698
- package/.claude/agents/sparc/pseudocode.md +519 -519
- package/.claude/agents/sparc/refinement.md +801 -801
- package/.claude/agents/sparc/specification.md +477 -477
- package/.claude/agents/specialized/mobile/spec-mobile-react-native.md +224 -224
- package/.claude/agents/specialized/spec-mobile-react-native.md +226 -226
- package/.claude/agents/sublinear/consensus-coordinator.md +337 -337
- package/.claude/agents/sublinear/matrix-optimizer.md +184 -184
- package/.claude/agents/sublinear/pagerank-analyzer.md +298 -298
- package/.claude/agents/sublinear/performance-optimizer.md +367 -367
- package/.claude/agents/sublinear/trading-predictor.md +245 -245
- package/.claude/agents/swarm/adaptive-coordinator.md +1126 -1126
- package/.claude/agents/swarm/hierarchical-coordinator.md +709 -709
- package/.claude/agents/swarm/mesh-coordinator.md +962 -962
- package/.claude/agents/templates/automation-smart-agent.md +204 -204
- package/.claude/agents/templates/base-template-generator.md +289 -289
- package/.claude/agents/templates/coordinator-swarm-init.md +89 -89
- package/.claude/agents/templates/github-pr-manager.md +176 -176
- package/.claude/agents/templates/implementer-sparc-coder.md +258 -258
- package/.claude/agents/templates/memory-coordinator.md +186 -186
- package/.claude/agents/templates/orchestrator-task.md +138 -138
- package/.claude/agents/templates/performance-analyzer.md +198 -198
- package/.claude/agents/templates/sparc-coordinator.md +513 -513
- package/.claude/agents/testing/production-validator.md +394 -394
- package/.claude/agents/testing/tdd-london-swarm.md +243 -243
- package/.claude/agents/v3/aidefence-guardian.md +282 -282
- package/.claude/agents/v3/claims-authorizer.md +208 -208
- package/.claude/agents/v3/collective-intelligence-coordinator.md +993 -993
- package/.claude/agents/v3/ddd-domain-expert.md +220 -220
- package/.claude/agents/v3/injection-analyst.md +236 -236
- package/.claude/agents/v3/performance-engineer.md +1233 -1233
- package/.claude/agents/v3/pii-detector.md +151 -151
- package/.claude/agents/v3/reasoningbank-learner.md +213 -213
- package/.claude/agents/v3/security-architect-aidefence.md +410 -410
- package/.claude/agents/v3/security-architect.md +867 -867
- package/.claude/agents/v3/swarm-memory-manager.md +157 -157
- package/.claude/agents/v3/v3-integration-architect.md +205 -205
- package/.claude/commands/agents/README.md +50 -50
- package/.claude/commands/agents/agent-capabilities.md +140 -140
- package/.claude/commands/agents/agent-coordination.md +28 -28
- package/.claude/commands/agents/agent-spawning.md +28 -28
- package/.claude/commands/agents/agent-types.md +216 -216
- package/.claude/commands/agents/health.md +139 -139
- package/.claude/commands/agents/list.md +100 -100
- package/.claude/commands/agents/logs.md +130 -130
- package/.claude/commands/agents/metrics.md +122 -122
- package/.claude/commands/agents/pool.md +127 -127
- package/.claude/commands/agents/spawn.md +140 -140
- package/.claude/commands/agents/status.md +115 -115
- package/.claude/commands/agents/stop.md +102 -102
- package/.claude/commands/analysis/COMMAND_COMPLIANCE_REPORT.md +53 -53
- package/.claude/commands/analysis/README.md +9 -9
- package/.claude/commands/analysis/bottleneck-detect.md +162 -162
- package/.claude/commands/analysis/performance-bottlenecks.md +58 -58
- package/.claude/commands/analysis/performance-report.md +25 -25
- package/.claude/commands/analysis/token-efficiency.md +44 -44
- package/.claude/commands/analysis/token-usage.md +25 -25
- package/.claude/commands/automation/README.md +9 -9
- package/.claude/commands/automation/auto-agent.md +122 -122
- package/.claude/commands/automation/self-healing.md +105 -105
- package/.claude/commands/automation/session-memory.md +89 -89
- package/.claude/commands/automation/smart-agents.md +72 -72
- package/.claude/commands/automation/smart-spawn.md +25 -25
- package/.claude/commands/automation/workflow-select.md +25 -25
- package/.claude/commands/claude-flow-help.md +103 -103
- package/.claude/commands/claude-flow-memory.md +107 -107
- package/.claude/commands/claude-flow-swarm.md +205 -205
- package/.claude/commands/coordination/README.md +9 -9
- package/.claude/commands/coordination/agent-spawn.md +25 -25
- package/.claude/commands/coordination/init.md +44 -44
- package/.claude/commands/coordination/orchestrate.md +43 -43
- package/.claude/commands/coordination/spawn.md +45 -45
- package/.claude/commands/coordination/swarm-init.md +85 -85
- package/.claude/commands/coordination/task-orchestrate.md +25 -25
- package/.claude/commands/github/README.md +11 -11
- package/.claude/commands/github/code-review-swarm.md +513 -513
- package/.claude/commands/github/code-review.md +25 -25
- package/.claude/commands/github/github-modes.md +146 -146
- package/.claude/commands/github/github-swarm.md +121 -121
- package/.claude/commands/github/issue-tracker.md +291 -291
- package/.claude/commands/github/issue-triage.md +25 -25
- package/.claude/commands/github/multi-repo-swarm.md +518 -518
- package/.claude/commands/github/pr-enhance.md +26 -26
- package/.claude/commands/github/pr-manager.md +169 -169
- package/.claude/commands/github/project-board-sync.md +470 -470
- package/.claude/commands/github/release-manager.md +339 -339
- package/.claude/commands/github/release-swarm.md +543 -543
- package/.claude/commands/github/repo-analyze.md +25 -25
- package/.claude/commands/github/repo-architect.md +366 -366
- package/.claude/commands/github/swarm-issue.md +484 -484
- package/.claude/commands/github/swarm-pr.md +287 -287
- package/.claude/commands/github/sync-coordinator.md +302 -302
- package/.claude/commands/github/workflow-automation.md +441 -441
- package/.claude/commands/hive-mind/README.md +17 -17
- package/.claude/commands/hive-mind/hive-mind-consensus.md +8 -8
- package/.claude/commands/hive-mind/hive-mind-init.md +18 -18
- package/.claude/commands/hive-mind/hive-mind-memory.md +8 -8
- package/.claude/commands/hive-mind/hive-mind-metrics.md +8 -8
- package/.claude/commands/hive-mind/hive-mind-resume.md +8 -8
- package/.claude/commands/hive-mind/hive-mind-sessions.md +8 -8
- package/.claude/commands/hive-mind/hive-mind-spawn.md +21 -21
- package/.claude/commands/hive-mind/hive-mind-status.md +8 -8
- package/.claude/commands/hive-mind/hive-mind-stop.md +8 -8
- package/.claude/commands/hive-mind/hive-mind-wizard.md +8 -8
- package/.claude/commands/hive-mind/hive-mind.md +27 -27
- package/.claude/commands/hooks/README.md +11 -11
- package/.claude/commands/hooks/overview.md +57 -57
- package/.claude/commands/hooks/post-edit.md +117 -117
- package/.claude/commands/hooks/post-task.md +112 -112
- package/.claude/commands/hooks/pre-edit.md +113 -113
- package/.claude/commands/hooks/pre-task.md +111 -111
- package/.claude/commands/hooks/session-end.md +118 -118
- package/.claude/commands/hooks/setup.md +102 -102
- package/.claude/commands/memory/README.md +9 -9
- package/.claude/commands/memory/memory-persist.md +25 -25
- package/.claude/commands/memory/memory-search.md +25 -25
- package/.claude/commands/memory/memory-usage.md +25 -25
- package/.claude/commands/memory/neural.md +47 -47
- package/.claude/commands/monitoring/README.md +9 -9
- package/.claude/commands/monitoring/agent-metrics.md +25 -25
- package/.claude/commands/monitoring/agents.md +44 -44
- package/.claude/commands/monitoring/real-time-view.md +25 -25
- package/.claude/commands/monitoring/status.md +46 -46
- package/.claude/commands/monitoring/swarm-monitor.md +25 -25
- package/.claude/commands/optimization/README.md +9 -9
- package/.claude/commands/optimization/auto-topology.md +61 -61
- package/.claude/commands/optimization/cache-manage.md +25 -25
- package/.claude/commands/optimization/parallel-execute.md +25 -25
- package/.claude/commands/optimization/parallel-execution.md +49 -49
- package/.claude/commands/optimization/topology-optimize.md +25 -25
- package/.claude/commands/pair/README.md +260 -260
- package/.claude/commands/pair/commands.md +545 -545
- package/.claude/commands/pair/config.md +509 -509
- package/.claude/commands/pair/examples.md +511 -511
- package/.claude/commands/pair/modes.md +347 -347
- package/.claude/commands/pair/session.md +406 -406
- package/.claude/commands/pair/start.md +208 -208
- package/.claude/commands/sparc/analyzer.md +51 -51
- package/.claude/commands/sparc/architect.md +53 -53
- package/.claude/commands/sparc/ask.md +97 -97
- package/.claude/commands/sparc/batch-executor.md +54 -54
- package/.claude/commands/sparc/code.md +89 -89
- package/.claude/commands/sparc/coder.md +54 -54
- package/.claude/commands/sparc/debug.md +83 -83
- package/.claude/commands/sparc/debugger.md +54 -54
- package/.claude/commands/sparc/designer.md +53 -53
- package/.claude/commands/sparc/devops.md +109 -109
- package/.claude/commands/sparc/docs-writer.md +80 -80
- package/.claude/commands/sparc/documenter.md +54 -54
- package/.claude/commands/sparc/innovator.md +54 -54
- package/.claude/commands/sparc/integration.md +83 -83
- package/.claude/commands/sparc/mcp.md +117 -117
- package/.claude/commands/sparc/memory-manager.md +54 -54
- package/.claude/commands/sparc/optimizer.md +54 -54
- package/.claude/commands/sparc/orchestrator.md +131 -131
- package/.claude/commands/sparc/post-deployment-monitoring-mode.md +83 -83
- package/.claude/commands/sparc/refinement-optimization-mode.md +83 -83
- package/.claude/commands/sparc/researcher.md +54 -54
- package/.claude/commands/sparc/reviewer.md +54 -54
- package/.claude/commands/sparc/security-review.md +80 -80
- package/.claude/commands/sparc/sparc-modes.md +174 -174
- package/.claude/commands/sparc/sparc.md +111 -111
- package/.claude/commands/sparc/spec-pseudocode.md +80 -80
- package/.claude/commands/sparc/supabase-admin.md +348 -348
- package/.claude/commands/sparc/swarm-coordinator.md +54 -54
- package/.claude/commands/sparc/tdd.md +54 -54
- package/.claude/commands/sparc/tester.md +54 -54
- package/.claude/commands/sparc/tutorial.md +79 -79
- package/.claude/commands/sparc/workflow-manager.md +54 -54
- package/.claude/commands/sparc.md +166 -166
- package/.claude/commands/stream-chain/pipeline.md +120 -120
- package/.claude/commands/stream-chain/run.md +69 -69
- package/.claude/commands/swarm/README.md +15 -15
- package/.claude/commands/swarm/analysis.md +95 -95
- package/.claude/commands/swarm/development.md +96 -96
- package/.claude/commands/swarm/examples.md +168 -168
- package/.claude/commands/swarm/maintenance.md +102 -102
- package/.claude/commands/swarm/optimization.md +117 -117
- package/.claude/commands/swarm/research.md +136 -136
- package/.claude/commands/swarm/swarm-analysis.md +8 -8
- package/.claude/commands/swarm/swarm-background.md +8 -8
- package/.claude/commands/swarm/swarm-init.md +19 -19
- package/.claude/commands/swarm/swarm-modes.md +8 -8
- package/.claude/commands/swarm/swarm-monitor.md +8 -8
- package/.claude/commands/swarm/swarm-spawn.md +19 -19
- package/.claude/commands/swarm/swarm-status.md +8 -8
- package/.claude/commands/swarm/swarm-strategies.md +8 -8
- package/.claude/commands/swarm/swarm.md +87 -87
- package/.claude/commands/swarm/testing.md +131 -131
- package/.claude/commands/training/README.md +9 -9
- package/.claude/commands/training/model-update.md +25 -25
- package/.claude/commands/training/neural-patterns.md +107 -107
- package/.claude/commands/training/neural-train.md +75 -75
- package/.claude/commands/training/pattern-learn.md +25 -25
- package/.claude/commands/training/specialization.md +62 -62
- package/.claude/commands/truth/start.md +142 -142
- package/.claude/commands/verify/check.md +49 -49
- package/.claude/commands/verify/start.md +127 -127
- package/.claude/commands/workflows/README.md +9 -9
- package/.claude/commands/workflows/development.md +77 -77
- package/.claude/commands/workflows/research.md +62 -62
- package/.claude/commands/workflows/workflow-create.md +25 -25
- package/.claude/commands/workflows/workflow-execute.md +25 -25
- package/.claude/commands/workflows/workflow-export.md +25 -25
- package/.claude/eval/human-relevance-frozen-v1.json +17 -17
- package/.claude/evolve-proof/generation-0.json +211 -211
- package/.claude/evolve-proof/real-generation-0.json +406 -406
- package/.claude/evolve-proof/real-generation-1.json +406 -406
- package/.claude/helpers/README.md +96 -96
- package/.claude/helpers/adr-compliance.sh +186 -186
- package/.claude/helpers/auto-commit.sh +178 -178
- package/.claude/helpers/auto-memory-hook.mjs +430 -430
- package/.claude/helpers/checkpoint-manager.sh +251 -251
- package/.claude/helpers/daemon-manager.sh +252 -252
- package/.claude/helpers/ddd-tracker.sh +144 -144
- package/.claude/helpers/github-safe.js +156 -156
- package/.claude/helpers/github-setup.sh +45 -45
- package/.claude/helpers/guidance-hook.sh +13 -13
- package/.claude/helpers/guidance-hooks.sh +102 -102
- package/.claude/helpers/health-monitor.sh +108 -108
- package/.claude/helpers/helpers.manifest.json +6 -6
- package/.claude/helpers/hook-handler.cjs +565 -565
- package/.claude/helpers/intelligence.cjs +1058 -1058
- package/.claude/helpers/learning-hooks.sh +329 -329
- package/.claude/helpers/learning-optimizer.sh +127 -127
- package/.claude/helpers/learning-service.mjs +1144 -1144
- package/.claude/helpers/memory.js +83 -83
- package/.claude/helpers/metrics-db.mjs +503 -503
- package/.claude/helpers/pattern-consolidator.sh +86 -86
- package/.claude/helpers/perf-worker.sh +160 -160
- package/.claude/helpers/post-commit +16 -16
- package/.claude/helpers/pre-commit +26 -26
- package/.claude/helpers/quick-start.sh +19 -19
- package/.claude/helpers/router.js +105 -105
- package/.claude/helpers/security-scanner.sh +127 -127
- package/.claude/helpers/session.js +157 -157
- package/.claude/helpers/setup-mcp.sh +18 -18
- package/.claude/helpers/standard-checkpoint-hooks.sh +189 -189
- package/.claude/helpers/statusline-hook.sh +21 -21
- package/.claude/helpers/statusline.cjs +1060 -1060
- package/.claude/helpers/statusline.js +340 -340
- package/.claude/helpers/swarm-comms.sh +353 -353
- package/.claude/helpers/swarm-hooks.sh +761 -761
- package/.claude/helpers/swarm-monitor.sh +210 -210
- package/.claude/helpers/sync-v3-metrics.sh +245 -245
- package/.claude/helpers/update-v3-progress.sh +165 -165
- package/.claude/helpers/v3-quick-status.sh +57 -57
- package/.claude/helpers/v3.sh +110 -110
- package/.claude/helpers/validate-v3-config.sh +215 -215
- package/.claude/helpers/worker-manager.sh +170 -170
- package/.claude/proven-config.manifest.json +37 -37
- package/.claude/proven-config.signed.json +41 -41
- package/.claude/settings.json +182 -182
- package/.claude/skills/agentdb-advanced/SKILL.md +550 -550
- package/.claude/skills/agentdb-learning/SKILL.md +545 -545
- package/.claude/skills/agentdb-memory-patterns/SKILL.md +339 -339
- package/.claude/skills/agentdb-optimization/SKILL.md +509 -509
- package/.claude/skills/agentdb-vector-search/SKILL.md +339 -339
- package/.claude/skills/browser/SKILL.md +204 -204
- package/.claude/skills/dual-mode/README.md +71 -71
- package/.claude/skills/dual-mode/dual-collect.md +103 -103
- package/.claude/skills/dual-mode/dual-coordinate.md +85 -85
- package/.claude/skills/dual-mode/dual-spawn.md +81 -81
- package/.claude/skills/flow-nexus-neural/SKILL.md +727 -727
- package/.claude/skills/flow-nexus-platform/SKILL.md +1154 -1154
- package/.claude/skills/flow-nexus-swarm/SKILL.md +604 -604
- package/.claude/skills/github-code-review/SKILL.md +1125 -1125
- package/.claude/skills/github-multi-repo/SKILL.md +862 -862
- package/.claude/skills/github-project-management/SKILL.md +1262 -1262
- package/.claude/skills/github-release-management/SKILL.md +1064 -1064
- package/.claude/skills/github-workflow-automation/SKILL.md +1047 -1047
- package/.claude/skills/hooks-automation/SKILL.md +1201 -1201
- package/.claude/skills/pair-programming/SKILL.md +1202 -1202
- package/.claude/skills/reasoningbank-agentdb/SKILL.md +446 -446
- package/.claude/skills/reasoningbank-intelligence/SKILL.md +201 -201
- package/.claude/skills/skill-builder/SKILL.md +910 -910
- package/.claude/skills/sparc-methodology/SKILL.md +1106 -1106
- package/.claude/skills/stream-chain/SKILL.md +560 -560
- package/.claude/skills/swarm-advanced/SKILL.md +970 -970
- package/.claude/skills/swarm-orchestration/SKILL.md +179 -179
- package/.claude/skills/v3-cli-modernization/SKILL.md +871 -871
- package/.claude/skills/v3-core-implementation/SKILL.md +796 -796
- package/.claude/skills/v3-ddd-architecture/SKILL.md +441 -441
- package/.claude/skills/v3-integration-deep/SKILL.md +240 -240
- package/.claude/skills/v3-mcp-optimization/SKILL.md +776 -776
- package/.claude/skills/v3-memory-unification/SKILL.md +173 -173
- package/.claude/skills/v3-performance-optimization/SKILL.md +389 -389
- package/.claude/skills/v3-security-overhaul/SKILL.md +81 -81
- package/.claude/skills/v3-swarm-coordination/SKILL.md +339 -339
- package/.claude/skills/verification-quality/SKILL.md +691 -691
- package/README.md +419 -419
- package/bin/cli.js +314 -314
- package/bin/mcp-server.js +224 -224
- package/bin/preinstall.cjs +2 -2
- package/catalog-manifest.json +2 -2
- package/dist/src/autopilot-state.js +24 -7
- package/dist/src/benchmarks/gaia-critic.js +24 -24
- package/dist/src/business-pods/bbs-budget-tracker.js +53 -53
- package/dist/src/commands/completions.js +409 -409
- package/dist/src/commands/daemon.js +44 -44
- package/dist/src/commands/embeddings.js +26 -26
- package/dist/src/commands/hive-mind.js +97 -97
- package/dist/src/commands/hooks.js +31 -10
- package/dist/src/commands/init.js +202 -34
- package/dist/src/commands/memory.js +12 -1
- package/dist/src/commands/ruvector/backup.js +23 -23
- package/dist/src/commands/ruvector/benchmark.js +31 -31
- package/dist/src/commands/ruvector/import.js +14 -14
- package/dist/src/commands/ruvector/init.js +115 -115
- package/dist/src/commands/ruvector/migrate.js +99 -99
- package/dist/src/commands/ruvector/optimize.js +51 -51
- package/dist/src/commands/ruvector/setup.js +624 -624
- package/dist/src/commands/ruvector/status.js +38 -38
- package/dist/src/config/proven-config.js +2 -2
- package/dist/src/funnel/disclosure.js +13 -2
- package/dist/src/funnel/messages.d.ts +12 -10
- package/dist/src/funnel/messages.js +83 -11
- package/dist/src/init/claudemd-generator.js +231 -231
- package/dist/src/init/executor.js +453 -453
- package/dist/src/init/helper-signing.js +2 -2
- package/dist/src/init/helpers-generator.js +751 -751
- package/dist/src/init/statusline-generator.js +24 -24
- package/dist/src/mcp-tools/agentdb-tools.js +15 -15
- package/dist/src/mcp-tools/browser-intent-tools.js +19 -19
- package/dist/src/mcp-tools/browser-tools.js +8 -0
- package/dist/src/mcp-tools/hooks-tools.js +21 -0
- package/dist/src/mcp-tools/memory-tools.js +4 -3
- package/dist/src/memory/graph-edge-writer.js +22 -22
- package/dist/src/memory/memory-bridge.js +192 -123
- package/dist/src/memory/memory-initializer.js +407 -407
- package/dist/src/memory/rabitq-index.js +5 -5
- package/dist/src/parser.js +25 -9
- package/dist/src/proxy/verify.js +2 -2
- package/dist/src/runtime/headless.js +28 -28
- package/dist/src/services/distill-tuning.js +7 -7
- package/dist/src/services/headless-worker-executor.js +84 -84
- package/dist/src/services/memory-distillation.js +4 -4
- package/dist/src/services/worker-daemon.js +7 -4
- package/dist/src/transfer/deploy-seraphine.js +23 -23
- package/package.json +137 -137
- package/plugins/ruflo-metaharness/.claude-plugin/plugin.json +32 -32
- package/plugins/ruflo-metaharness/README.md +72 -72
- package/plugins/ruflo-metaharness/agents/metaharness-architect.md +58 -58
- package/plugins/ruflo-metaharness/commands/ruflo-metaharness.md +48 -48
- package/plugins/ruflo-metaharness/scripts/_darwin.mjs +210 -210
- package/plugins/ruflo-metaharness/scripts/_harness.mjs +330 -330
- package/plugins/ruflo-metaharness/scripts/_invoke.mjs +231 -231
- package/plugins/ruflo-metaharness/scripts/_redblue.mjs +143 -143
- package/plugins/ruflo-metaharness/scripts/_similarity.mjs +161 -161
- package/plugins/ruflo-metaharness/scripts/_spike-similarity.mjs +223 -223
- package/plugins/ruflo-metaharness/scripts/audit-list.mjs +158 -158
- package/plugins/ruflo-metaharness/scripts/audit-trend.mjs +272 -272
- package/plugins/ruflo-metaharness/scripts/bench-parse-mcp-scan.mjs +146 -146
- package/plugins/ruflo-metaharness/scripts/bench-recordpair-overhead.mjs +186 -186
- package/plugins/ruflo-metaharness/scripts/bench-similarity.mjs +177 -177
- package/plugins/ruflo-metaharness/scripts/bench.mjs +95 -95
- package/plugins/ruflo-metaharness/scripts/drift-from-history.mjs +363 -363
- package/plugins/ruflo-metaharness/scripts/evolve.mjs +404 -404
- package/plugins/ruflo-metaharness/scripts/genome.mjs +80 -80
- package/plugins/ruflo-metaharness/scripts/gepa.mjs +153 -153
- package/plugins/ruflo-metaharness/scripts/learn.mjs +127 -127
- package/plugins/ruflo-metaharness/scripts/mcp-scan.mjs +111 -111
- package/plugins/ruflo-metaharness/scripts/mint.mjs +126 -126
- package/plugins/ruflo-metaharness/scripts/oia-audit.mjs +228 -228
- package/plugins/ruflo-metaharness/scripts/redblue.mjs +286 -286
- package/plugins/ruflo-metaharness/scripts/router-parallel-analyze.mjs +250 -250
- package/plugins/ruflo-metaharness/scripts/score.mjs +92 -92
- package/plugins/ruflo-metaharness/scripts/security-bench.mjs +174 -174
- package/plugins/ruflo-metaharness/scripts/similarity.mjs +158 -158
- package/plugins/ruflo-metaharness/scripts/smoke.sh +2356 -2356
- package/plugins/ruflo-metaharness/scripts/test-graceful-degradation.mjs +165 -165
- package/plugins/ruflo-metaharness/scripts/test-mcp-tools.mjs +472 -472
- package/plugins/ruflo-metaharness/scripts/test-parallel-pipeline.mjs +204 -204
- package/plugins/ruflo-metaharness/scripts/test-pipeline-roundtrip.mjs +586 -586
- package/plugins/ruflo-metaharness/scripts/test-similarity.mjs +334 -334
- package/plugins/ruflo-metaharness/scripts/test-with-openrouter.mjs +229 -229
- package/plugins/ruflo-metaharness/scripts/threat-model.mjs +59 -59
- package/plugins/ruflo-metaharness/skills/harness-bench/SKILL.md +64 -64
- package/plugins/ruflo-metaharness/skills/harness-drift-from-history/SKILL.md +65 -65
- package/plugins/ruflo-metaharness/skills/harness-evolve/SKILL.md +131 -131
- package/plugins/ruflo-metaharness/skills/harness-genome/SKILL.md +54 -54
- package/plugins/ruflo-metaharness/skills/harness-gepa/SKILL.md +65 -65
- package/plugins/ruflo-metaharness/skills/harness-learn/SKILL.md +65 -65
- package/plugins/ruflo-metaharness/skills/harness-mcp-scan/SKILL.md +49 -49
- package/plugins/ruflo-metaharness/skills/harness-mint/SKILL.md +72 -72
- package/plugins/ruflo-metaharness/skills/harness-oia-audit/SKILL.md +79 -79
- package/plugins/ruflo-metaharness/skills/harness-score/SKILL.md +66 -66
- package/plugins/ruflo-metaharness/skills/harness-security-bench/SKILL.md +101 -101
- package/plugins/ruflo-metaharness/skills/harness-similarity/SKILL.md +67 -67
- package/plugins/ruflo-metaharness/skills/harness-threat-model/SKILL.md +41 -41
- package/scripts/postinstall.cjs +153 -153
|
@@ -1,65 +1,65 @@
|
|
|
1
|
-
---
|
|
2
|
-
name: harness-gepa
|
|
3
|
-
description: Inspect and audit GEPA genomes via the `@metaharness/darwin/gepa` library entry (darwin 0.8.0) — load/validate a genome (default: the shipped cand-6 promotion), render the system prompt a genome compiles to, or classify failure modes in a run transcript. The `gepaOptimize` loop itself is library-only (bring your own evaluator) and not surfaced here — use `harness-evolve` for sandbox-scored evolution. Degrades gracefully when @metaharness/darwin is absent.
|
|
4
|
-
argument-hint: "--op genome|validate|render|analyze [--path <genome.json>] [--transcript <t.json>] [--alert-on-invalid]"
|
|
5
|
-
allowed-tools: Bash
|
|
6
|
-
---
|
|
7
|
-
|
|
8
|
-
Surfaces the GEPA (genetic-evolution prompt-adaptation) *library* exports
|
|
9
|
-
from `@metaharness/darwin/gepa`. Unlike the other skills in this plugin
|
|
10
|
-
there is no CLI binary behind this — the script dynamic-imports the library
|
|
11
|
-
(local resolution first, versioned cache install as fallback) and calls the
|
|
12
|
-
subprocess-safe subset.
|
|
13
|
-
|
|
14
|
-
## When to use
|
|
15
|
-
|
|
16
|
-
- **Adopting an evolved policy**: `--op render` shows the actual system
|
|
17
|
-
prompt a genome compiles to — read THAT, not the raw JSON, before
|
|
18
|
-
wiring a genome into a harness.
|
|
19
|
-
- **Auditing a promotion**: `--op genome` loads + validates the shipped
|
|
20
|
-
cand-6 genome (first holdout-confirmed cheap-tier promotion; provenance
|
|
21
|
-
ships in the package) or any genome file you point at.
|
|
22
|
-
- **CI gate on genome edits**: `--op validate --alert-on-invalid` exits 1
|
|
23
|
-
on structural errors.
|
|
24
|
-
- **Debugging a bad run**: `--op analyze --transcript run.json` classifies
|
|
25
|
-
failure modes (GEPA's failure-class taxonomy) from a transcript array.
|
|
26
|
-
|
|
27
|
-
## What is deliberately NOT here
|
|
28
|
-
|
|
29
|
-
`gepaOptimize` — the optimization loop takes an in-process
|
|
30
|
-
`evaluate(candidate)` callback ("bring your own evaluator") that cannot
|
|
31
|
-
cross a subprocess boundary. Two supported paths instead:
|
|
32
|
-
|
|
33
|
-
1. **Library consumers**: `import { gepaOptimize, loadCand6Genome } from '@metaharness/darwin/gepa'`
|
|
34
|
-
2. **Sandbox-scored evolution**: `harness-evolve` (darwin CLI `evolve`),
|
|
35
|
-
which pairs GEPA with its own sandbox evaluators.
|
|
36
|
-
|
|
37
|
-
## Algorithm
|
|
38
|
-
|
|
39
|
-
Implementation: [`scripts/gepa.mjs`](../../scripts/gepa.mjs).
|
|
40
|
-
|
|
41
|
-
1. `import('@metaharness/darwin/gepa')`; on MODULE_NOT_FOUND fall back to a
|
|
42
|
-
one-time `npm install --prefix ~/.ruflo/darwin-cache-0.8.0` and import
|
|
43
|
-
the cached `dist/gepa/index.js` (versioned dir → pin bumps invalidate).
|
|
44
|
-
2. Dispatch `--op`:
|
|
45
|
-
- `genome` → `loadGenome(fs, path)` or `loadCand6Genome()` + `validateGenome`
|
|
46
|
-
- `validate` → `validateGenome(rawJson)` (raw parse so broken files reach
|
|
47
|
-
the validator instead of throwing in the loader)
|
|
48
|
-
- `render` → `buildSystemFromGenome(genome, ext?, glob?)`
|
|
49
|
-
- `analyze` → `analyzeTranscript(entries)`
|
|
50
|
-
3. Emit one JSON object; exit 0 (or 1 under `--alert-on-invalid`, 2 on bad input).
|
|
51
|
-
|
|
52
|
-
## Examples
|
|
53
|
-
|
|
54
|
-
```bash
|
|
55
|
-
node scripts/gepa.mjs --op genome # cand-6 + validation
|
|
56
|
-
node scripts/gepa.mjs --op render | jq -r .system # what does cand-6 SAY?
|
|
57
|
-
node scripts/gepa.mjs --op validate --path my-genome.json --alert-on-invalid
|
|
58
|
-
node scripts/gepa.mjs --op analyze --transcript run.json
|
|
59
|
-
```
|
|
60
|
-
|
|
61
|
-
## Exit codes
|
|
62
|
-
|
|
63
|
-
- `0` — op completed (or degraded — darwin not installable)
|
|
64
|
-
- `1` — `--alert-on-invalid` and validation found errors
|
|
65
|
-
- `2` — config error (unknown op, missing/broken input file)
|
|
1
|
+
---
|
|
2
|
+
name: harness-gepa
|
|
3
|
+
description: Inspect and audit GEPA genomes via the `@metaharness/darwin/gepa` library entry (darwin 0.8.0) — load/validate a genome (default: the shipped cand-6 promotion), render the system prompt a genome compiles to, or classify failure modes in a run transcript. The `gepaOptimize` loop itself is library-only (bring your own evaluator) and not surfaced here — use `harness-evolve` for sandbox-scored evolution. Degrades gracefully when @metaharness/darwin is absent.
|
|
4
|
+
argument-hint: "--op genome|validate|render|analyze [--path <genome.json>] [--transcript <t.json>] [--alert-on-invalid]"
|
|
5
|
+
allowed-tools: Bash
|
|
6
|
+
---
|
|
7
|
+
|
|
8
|
+
Surfaces the GEPA (genetic-evolution prompt-adaptation) *library* exports
|
|
9
|
+
from `@metaharness/darwin/gepa`. Unlike the other skills in this plugin
|
|
10
|
+
there is no CLI binary behind this — the script dynamic-imports the library
|
|
11
|
+
(local resolution first, versioned cache install as fallback) and calls the
|
|
12
|
+
subprocess-safe subset.
|
|
13
|
+
|
|
14
|
+
## When to use
|
|
15
|
+
|
|
16
|
+
- **Adopting an evolved policy**: `--op render` shows the actual system
|
|
17
|
+
prompt a genome compiles to — read THAT, not the raw JSON, before
|
|
18
|
+
wiring a genome into a harness.
|
|
19
|
+
- **Auditing a promotion**: `--op genome` loads + validates the shipped
|
|
20
|
+
cand-6 genome (first holdout-confirmed cheap-tier promotion; provenance
|
|
21
|
+
ships in the package) or any genome file you point at.
|
|
22
|
+
- **CI gate on genome edits**: `--op validate --alert-on-invalid` exits 1
|
|
23
|
+
on structural errors.
|
|
24
|
+
- **Debugging a bad run**: `--op analyze --transcript run.json` classifies
|
|
25
|
+
failure modes (GEPA's failure-class taxonomy) from a transcript array.
|
|
26
|
+
|
|
27
|
+
## What is deliberately NOT here
|
|
28
|
+
|
|
29
|
+
`gepaOptimize` — the optimization loop takes an in-process
|
|
30
|
+
`evaluate(candidate)` callback ("bring your own evaluator") that cannot
|
|
31
|
+
cross a subprocess boundary. Two supported paths instead:
|
|
32
|
+
|
|
33
|
+
1. **Library consumers**: `import { gepaOptimize, loadCand6Genome } from '@metaharness/darwin/gepa'`
|
|
34
|
+
2. **Sandbox-scored evolution**: `harness-evolve` (darwin CLI `evolve`),
|
|
35
|
+
which pairs GEPA with its own sandbox evaluators.
|
|
36
|
+
|
|
37
|
+
## Algorithm
|
|
38
|
+
|
|
39
|
+
Implementation: [`scripts/gepa.mjs`](../../scripts/gepa.mjs).
|
|
40
|
+
|
|
41
|
+
1. `import('@metaharness/darwin/gepa')`; on MODULE_NOT_FOUND fall back to a
|
|
42
|
+
one-time `npm install --prefix ~/.ruflo/darwin-cache-0.8.0` and import
|
|
43
|
+
the cached `dist/gepa/index.js` (versioned dir → pin bumps invalidate).
|
|
44
|
+
2. Dispatch `--op`:
|
|
45
|
+
- `genome` → `loadGenome(fs, path)` or `loadCand6Genome()` + `validateGenome`
|
|
46
|
+
- `validate` → `validateGenome(rawJson)` (raw parse so broken files reach
|
|
47
|
+
the validator instead of throwing in the loader)
|
|
48
|
+
- `render` → `buildSystemFromGenome(genome, ext?, glob?)`
|
|
49
|
+
- `analyze` → `analyzeTranscript(entries)`
|
|
50
|
+
3. Emit one JSON object; exit 0 (or 1 under `--alert-on-invalid`, 2 on bad input).
|
|
51
|
+
|
|
52
|
+
## Examples
|
|
53
|
+
|
|
54
|
+
```bash
|
|
55
|
+
node scripts/gepa.mjs --op genome # cand-6 + validation
|
|
56
|
+
node scripts/gepa.mjs --op render | jq -r .system # what does cand-6 SAY?
|
|
57
|
+
node scripts/gepa.mjs --op validate --path my-genome.json --alert-on-invalid
|
|
58
|
+
node scripts/gepa.mjs --op analyze --transcript run.json
|
|
59
|
+
```
|
|
60
|
+
|
|
61
|
+
## Exit codes
|
|
62
|
+
|
|
63
|
+
- `0` — op completed (or degraded — darwin not installable)
|
|
64
|
+
- `1` — `--alert-on-invalid` and validation found errors
|
|
65
|
+
- `2` — config error (unknown op, missing/broken input file)
|
|
@@ -1,65 +1,65 @@
|
|
|
1
|
-
---
|
|
2
|
-
name: harness-learn
|
|
3
|
-
description: Run a GEPA learning cycle via `metaharness learn` (upstream ADR-235, metaharness@0.3.0) — optimizes a harness genome against a SWE-bench-style slice manifest. $0 dry-run by default; `--run` is the explicit spend opt-in. Requires a metaharness repo checkout (`--repo` or $METAHARNESS_REPO) — without one it reports `checkout-required` with clone instructions. Degrades gracefully when metaharness is absent.
|
|
4
|
-
argument-hint: "--host <h> --model <m> --slice <manifest> [--repo <checkout>] [--run] [--alert-on-fail]"
|
|
5
|
-
allowed-tools: Bash
|
|
6
|
-
---
|
|
7
|
-
|
|
8
|
-
Surfaces `metaharness learn` — the upstream GEPA learning harness that
|
|
9
|
-
evolves harness policy genomes against a scored task corpus instead of
|
|
10
|
-
hand-editing prompts. Candidates are scored on held-out slices and only
|
|
11
|
-
measured winners promote (the shipped cand-6 genome is the first such
|
|
12
|
-
promotion: holdout gold 2/12 → 3/12, zero regressions).
|
|
13
|
-
|
|
14
|
-
## When to use
|
|
15
|
-
|
|
16
|
-
- A harness's policy prompt underperforms on a task family and you want a
|
|
17
|
-
measured improvement loop rather than manual prompt iteration.
|
|
18
|
-
- Pricing a learning run before committing spend — the default dry-run
|
|
19
|
-
resolves the slice manifest and reports cost without any model calls.
|
|
20
|
-
- After a learn run promotes a genome: pair with `harness-gepa --op render`
|
|
21
|
-
to inspect what the promoted policy actually says.
|
|
22
|
-
|
|
23
|
-
## Preconditions (upstream design)
|
|
24
|
-
|
|
25
|
-
The learning harness (GEPA + SWE-bench + Docker) is too heavy for the npm
|
|
26
|
-
package, so `learn` needs a local clone:
|
|
27
|
-
|
|
28
|
-
```bash
|
|
29
|
-
git clone https://github.com/ruvnet/metaharness.git
|
|
30
|
-
node scripts/learn.mjs --repo ./metaharness --host claude-code --model haiku --slice slices/lite.json
|
|
31
|
-
```
|
|
32
|
-
|
|
33
|
-
Without a checkout the script emits `{status: "checkout-required"}` and
|
|
34
|
-
exits 0 — a precondition report, not an error (distinct from
|
|
35
|
-
`degraded: true`, which means the npm package itself is absent). The
|
|
36
|
-
managed-service path (gateway-side learn jobs, no checkout) is upstream's
|
|
37
|
-
ADR-235 follow-up and not available yet.
|
|
38
|
-
|
|
39
|
-
## Algorithm
|
|
40
|
-
|
|
41
|
-
Implementation: [`scripts/learn.mjs`](../../scripts/learn.mjs).
|
|
42
|
-
|
|
43
|
-
1. Validate `--repo` exists when given; export it as `$METAHARNESS_REPO`.
|
|
44
|
-
2. Invoke the pinned `metaharness` binary (`metaharness@~0.3.0`, local install
|
|
45
|
-
or one-time versioned cache — never `@latest`): `metaharness learn --host <h>
|
|
46
|
-
--model <m> --slice <s> [--run]` via `_harness.mjs` (graceful degradation,
|
|
47
|
-
hard timeout).
|
|
48
|
-
3. Default timeouts: 120s dry-run, 600s with `--run` — real runs on larger
|
|
49
|
-
slices need an explicit `--timeout-ms` matched to slice size × model cost.
|
|
50
|
-
4. Detect the checkout-required message → structured payload, exit 0.
|
|
51
|
-
5. Parse the trailing JSON report when upstream emits one; otherwise return
|
|
52
|
-
the raw report text under `rawReport`.
|
|
53
|
-
|
|
54
|
-
## Cost note
|
|
55
|
-
|
|
56
|
-
`--run` is the ONLY path that spends. Everything else — dry-run, checkout
|
|
57
|
-
probe, degraded path — is $0. The MCP tool (`metaharness_learn`) has a 120s
|
|
58
|
-
subprocess budget; run real learning cycles from a terminal via
|
|
59
|
-
`ruflo metaharness learn ... --run --timeout-ms <big>`.
|
|
60
|
-
|
|
61
|
-
## Exit codes
|
|
62
|
-
|
|
63
|
-
- `0` — report produced (or dry-run, checkout-required, degraded)
|
|
64
|
-
- `1` — `--alert-on-fail` and the learn run reported failure
|
|
65
|
-
- `2` — config error (bad `--repo` path)
|
|
1
|
+
---
|
|
2
|
+
name: harness-learn
|
|
3
|
+
description: Run a GEPA learning cycle via `metaharness learn` (upstream ADR-235, metaharness@0.3.0) — optimizes a harness genome against a SWE-bench-style slice manifest. $0 dry-run by default; `--run` is the explicit spend opt-in. Requires a metaharness repo checkout (`--repo` or $METAHARNESS_REPO) — without one it reports `checkout-required` with clone instructions. Degrades gracefully when metaharness is absent.
|
|
4
|
+
argument-hint: "--host <h> --model <m> --slice <manifest> [--repo <checkout>] [--run] [--alert-on-fail]"
|
|
5
|
+
allowed-tools: Bash
|
|
6
|
+
---
|
|
7
|
+
|
|
8
|
+
Surfaces `metaharness learn` — the upstream GEPA learning harness that
|
|
9
|
+
evolves harness policy genomes against a scored task corpus instead of
|
|
10
|
+
hand-editing prompts. Candidates are scored on held-out slices and only
|
|
11
|
+
measured winners promote (the shipped cand-6 genome is the first such
|
|
12
|
+
promotion: holdout gold 2/12 → 3/12, zero regressions).
|
|
13
|
+
|
|
14
|
+
## When to use
|
|
15
|
+
|
|
16
|
+
- A harness's policy prompt underperforms on a task family and you want a
|
|
17
|
+
measured improvement loop rather than manual prompt iteration.
|
|
18
|
+
- Pricing a learning run before committing spend — the default dry-run
|
|
19
|
+
resolves the slice manifest and reports cost without any model calls.
|
|
20
|
+
- After a learn run promotes a genome: pair with `harness-gepa --op render`
|
|
21
|
+
to inspect what the promoted policy actually says.
|
|
22
|
+
|
|
23
|
+
## Preconditions (upstream design)
|
|
24
|
+
|
|
25
|
+
The learning harness (GEPA + SWE-bench + Docker) is too heavy for the npm
|
|
26
|
+
package, so `learn` needs a local clone:
|
|
27
|
+
|
|
28
|
+
```bash
|
|
29
|
+
git clone https://github.com/ruvnet/metaharness.git
|
|
30
|
+
node scripts/learn.mjs --repo ./metaharness --host claude-code --model haiku --slice slices/lite.json
|
|
31
|
+
```
|
|
32
|
+
|
|
33
|
+
Without a checkout the script emits `{status: "checkout-required"}` and
|
|
34
|
+
exits 0 — a precondition report, not an error (distinct from
|
|
35
|
+
`degraded: true`, which means the npm package itself is absent). The
|
|
36
|
+
managed-service path (gateway-side learn jobs, no checkout) is upstream's
|
|
37
|
+
ADR-235 follow-up and not available yet.
|
|
38
|
+
|
|
39
|
+
## Algorithm
|
|
40
|
+
|
|
41
|
+
Implementation: [`scripts/learn.mjs`](../../scripts/learn.mjs).
|
|
42
|
+
|
|
43
|
+
1. Validate `--repo` exists when given; export it as `$METAHARNESS_REPO`.
|
|
44
|
+
2. Invoke the pinned `metaharness` binary (`metaharness@~0.3.0`, local install
|
|
45
|
+
or one-time versioned cache — never `@latest`): `metaharness learn --host <h>
|
|
46
|
+
--model <m> --slice <s> [--run]` via `_harness.mjs` (graceful degradation,
|
|
47
|
+
hard timeout).
|
|
48
|
+
3. Default timeouts: 120s dry-run, 600s with `--run` — real runs on larger
|
|
49
|
+
slices need an explicit `--timeout-ms` matched to slice size × model cost.
|
|
50
|
+
4. Detect the checkout-required message → structured payload, exit 0.
|
|
51
|
+
5. Parse the trailing JSON report when upstream emits one; otherwise return
|
|
52
|
+
the raw report text under `rawReport`.
|
|
53
|
+
|
|
54
|
+
## Cost note
|
|
55
|
+
|
|
56
|
+
`--run` is the ONLY path that spends. Everything else — dry-run, checkout
|
|
57
|
+
probe, degraded path — is $0. The MCP tool (`metaharness_learn`) has a 120s
|
|
58
|
+
subprocess budget; run real learning cycles from a terminal via
|
|
59
|
+
`ruflo metaharness learn ... --run --timeout-ms <big>`.
|
|
60
|
+
|
|
61
|
+
## Exit codes
|
|
62
|
+
|
|
63
|
+
- `0` — report produced (or dry-run, checkout-required, degraded)
|
|
64
|
+
- `1` — `--alert-on-fail` and the learn run reported failure
|
|
65
|
+
- `2` — config error (bad `--repo` path)
|
|
@@ -1,49 +1,49 @@
|
|
|
1
|
-
---
|
|
2
|
-
name: harness-mcp-scan
|
|
3
|
-
description: Static security scan of a harness's declared MCP surface via `harness mcp-scan <path>`. Reads `.mcp/servers.json` + `.harness/claims.json`. Pure-read, no dispatch. Exits 1 on findings at or above `--fail-on` severity.
|
|
4
|
-
argument-hint: "[--path .] [--fail-on low|medium|high] [--format table|json]"
|
|
5
|
-
allowed-tools: Bash
|
|
6
|
-
---
|
|
7
|
-
|
|
8
|
-
Calls `harness mcp-scan` to enumerate every declared MCP server + tool
|
|
9
|
-
and flag policy / permission / dependency issues. Never executes any
|
|
10
|
-
tool; pure static analysis.
|
|
11
|
-
|
|
12
|
-
## Algorithm
|
|
13
|
-
|
|
14
|
-
Implementation: [`scripts/mcp-scan.mjs`](../../scripts/mcp-scan.mjs).
|
|
15
|
-
|
|
16
|
-
1. Invoke the pinned `harness` binary (`metaharness@~0.3.0`, resolved from a
|
|
17
|
-
local install or the one-time `~/.ruflo/metaharness-cache-<pin>` cache —
|
|
18
|
-
never `@latest`): `harness mcp-scan <path> --json`.
|
|
19
|
-
2. Parse `findings[]` with `{ severity, id, server, tool, message }`.
|
|
20
|
-
3. `--fail-on <severity>`: exit 1 when any finding is at or above that
|
|
21
|
-
level. Default `high`.
|
|
22
|
-
4. Output JSON (default) or markdown table.
|
|
23
|
-
|
|
24
|
-
## Severity rank
|
|
25
|
-
|
|
26
|
-
| Severity | Rank |
|
|
27
|
-
|---|---:|
|
|
28
|
-
| low | 1 |
|
|
29
|
-
| medium | 2 |
|
|
30
|
-
| high | 3 |
|
|
31
|
-
|
|
32
|
-
`--fail-on high` (default) only fails on HIGH; `--fail-on medium` also
|
|
33
|
-
fails on MEDIUM; `--fail-on low` fails on any finding.
|
|
34
|
-
|
|
35
|
-
## CI integration
|
|
36
|
-
|
|
37
|
-
```yaml
|
|
38
|
-
- name: MCP static scan
|
|
39
|
-
run: node plugins/ruflo-metaharness/scripts/mcp-scan.mjs --fail-on high
|
|
40
|
-
```
|
|
41
|
-
|
|
42
|
-
The exit code is the only thing CI watches; the JSON output goes to
|
|
43
|
-
artifacts for human review.
|
|
44
|
-
|
|
45
|
-
## Graceful degradation
|
|
46
|
-
|
|
47
|
-
When `harness` binary is unavailable (no network, blocked registry),
|
|
48
|
-
emits structured `{ degraded: true, reason: 'metaharness-not-available' }`
|
|
49
|
-
and exits 0. Ruflo continues — ADR-150 architectural constraint.
|
|
1
|
+
---
|
|
2
|
+
name: harness-mcp-scan
|
|
3
|
+
description: Static security scan of a harness's declared MCP surface via `harness mcp-scan <path>`. Reads `.mcp/servers.json` + `.harness/claims.json`. Pure-read, no dispatch. Exits 1 on findings at or above `--fail-on` severity.
|
|
4
|
+
argument-hint: "[--path .] [--fail-on low|medium|high] [--format table|json]"
|
|
5
|
+
allowed-tools: Bash
|
|
6
|
+
---
|
|
7
|
+
|
|
8
|
+
Calls `harness mcp-scan` to enumerate every declared MCP server + tool
|
|
9
|
+
and flag policy / permission / dependency issues. Never executes any
|
|
10
|
+
tool; pure static analysis.
|
|
11
|
+
|
|
12
|
+
## Algorithm
|
|
13
|
+
|
|
14
|
+
Implementation: [`scripts/mcp-scan.mjs`](../../scripts/mcp-scan.mjs).
|
|
15
|
+
|
|
16
|
+
1. Invoke the pinned `harness` binary (`metaharness@~0.3.0`, resolved from a
|
|
17
|
+
local install or the one-time `~/.ruflo/metaharness-cache-<pin>` cache —
|
|
18
|
+
never `@latest`): `harness mcp-scan <path> --json`.
|
|
19
|
+
2. Parse `findings[]` with `{ severity, id, server, tool, message }`.
|
|
20
|
+
3. `--fail-on <severity>`: exit 1 when any finding is at or above that
|
|
21
|
+
level. Default `high`.
|
|
22
|
+
4. Output JSON (default) or markdown table.
|
|
23
|
+
|
|
24
|
+
## Severity rank
|
|
25
|
+
|
|
26
|
+
| Severity | Rank |
|
|
27
|
+
|---|---:|
|
|
28
|
+
| low | 1 |
|
|
29
|
+
| medium | 2 |
|
|
30
|
+
| high | 3 |
|
|
31
|
+
|
|
32
|
+
`--fail-on high` (default) only fails on HIGH; `--fail-on medium` also
|
|
33
|
+
fails on MEDIUM; `--fail-on low` fails on any finding.
|
|
34
|
+
|
|
35
|
+
## CI integration
|
|
36
|
+
|
|
37
|
+
```yaml
|
|
38
|
+
- name: MCP static scan
|
|
39
|
+
run: node plugins/ruflo-metaharness/scripts/mcp-scan.mjs --fail-on high
|
|
40
|
+
```
|
|
41
|
+
|
|
42
|
+
The exit code is the only thing CI watches; the JSON output goes to
|
|
43
|
+
artifacts for human review.
|
|
44
|
+
|
|
45
|
+
## Graceful degradation
|
|
46
|
+
|
|
47
|
+
When `harness` binary is unavailable (no network, blocked registry),
|
|
48
|
+
emits structured `{ degraded: true, reason: 'metaharness-not-available' }`
|
|
49
|
+
and exits 0. Ruflo continues — ADR-150 architectural constraint.
|
|
@@ -1,72 +1,72 @@
|
|
|
1
|
-
---
|
|
2
|
-
name: harness-mint
|
|
3
|
-
description: Scaffold a custom AI agent harness via `metaharness new <name> --template <id> --host <id>`. Defaults to DRY-RUN (no writes) unless --confirm is passed. Refuses to write to the calling repo root or anywhere inside it. Honors ADR-150 architectural constraint + ruflo's "destructive-action confirmation" pattern.
|
|
4
|
-
argument-hint: "--name <id> --template <vertical:coding|minimal|…> [--host claude-code|codex|…] [--target /abs/path] [--confirm] [--format table|json]"
|
|
5
|
-
allowed-tools: Bash
|
|
6
|
-
---
|
|
7
|
-
|
|
8
|
-
The one write-capable skill in the plugin. Every other skill is
|
|
9
|
-
pure-read. This one calls `metaharness new`, which writes a new
|
|
10
|
-
directory tree.
|
|
11
|
-
|
|
12
|
-
## Safety (load-bearing)
|
|
13
|
-
|
|
14
|
-
1. **Dry-run by default.** Without `--confirm`, the script prints what
|
|
15
|
-
it would do and exits 0 without touching disk.
|
|
16
|
-
2. **Refuses project root.** If `--target` resolves to the current
|
|
17
|
-
working directory OR any path inside it, the script errors out with
|
|
18
|
-
exit 2. Target must be an absolute path OUTSIDE the calling repo
|
|
19
|
-
(default is a fresh `/tmp/ruflo-mint-<ts>-<name>/` dir).
|
|
20
|
-
3. **Refuses existing target.** Won't overwrite — must scaffold into a
|
|
21
|
-
non-existent dir.
|
|
22
|
-
4. **Subprocess + 60s timeout.** No library import, no in-process
|
|
23
|
-
execution. The mint stays sandboxed from ruflo's runtime.
|
|
24
|
-
|
|
25
|
-
## Algorithm
|
|
26
|
-
|
|
27
|
-
Implementation: [`scripts/mint.mjs`](../../scripts/mint.mjs).
|
|
28
|
-
|
|
29
|
-
1. Validate `--name`, `--template`. Default `--host` to `claude-code`.
|
|
30
|
-
2. Resolve `--target` (default: temp dir).
|
|
31
|
-
3. Run safety checks (no project-root writes; target must not exist).
|
|
32
|
-
4. Without `--confirm`: emit dry-run plan, exit 0.
|
|
33
|
-
5. With `--confirm`: shell `npx metaharness new <name> --template <id>
|
|
34
|
-
--host <id> --target <abs> --yes`.
|
|
35
|
-
|
|
36
|
-
## Templates
|
|
37
|
-
|
|
38
|
-
`minimal`, `vertical:coding`, `vertical:devops`, `vertical:support`,
|
|
39
|
-
`vertical:legal`, `vertical:research`, `vertical:trading`, `vertical:health`,
|
|
40
|
-
`vertical:education`, `vertical:sales`, `vertical:business`,
|
|
41
|
-
`vertical:crm`, `vertical:marketing`, `vertical:advertising`,
|
|
42
|
-
`vertical:ai`, `vertical:agentics`, `vertical:ruview`, `vertical:gaming`,
|
|
43
|
-
`vertical:repo-maintainer`, `vertical:exotic`.
|
|
44
|
-
|
|
45
|
-
## Hosts
|
|
46
|
-
|
|
47
|
-
`claude-code`, `codex`, `pi-dev`, `hermes`, `openclaw`, `rvm`,
|
|
48
|
-
`copilot`, `opencode`, `github-actions`.
|
|
49
|
-
|
|
50
|
-
## Example dry-run
|
|
51
|
-
|
|
52
|
-
```
|
|
53
|
-
$ node scripts/mint.mjs --name my-harness --template vertical:coding --host claude-code
|
|
54
|
-
# harness-mint (dry-run)
|
|
55
|
-
|
|
56
|
-
- action: metaharness new
|
|
57
|
-
- name: my-harness
|
|
58
|
-
- template: vertical:coding
|
|
59
|
-
- host: claude-code
|
|
60
|
-
- target: /tmp/ruflo-mint-1718560000-my-harness
|
|
61
|
-
- confirm: false
|
|
62
|
-
- willWrite: false
|
|
63
|
-
|
|
64
|
-
Re-run with `--confirm` to actually scaffold.
|
|
65
|
-
```
|
|
66
|
-
|
|
67
|
-
## Why dry-run by default
|
|
68
|
-
|
|
69
|
-
Ruflo's behavioral rules say "executing actions with care" — destructive
|
|
70
|
-
or repo-touching actions need confirmation. The dry-run output makes the
|
|
71
|
-
WHAT visible before the WHEN. A human sees `target`, decides, then
|
|
72
|
-
adds `--confirm` if happy.
|
|
1
|
+
---
|
|
2
|
+
name: harness-mint
|
|
3
|
+
description: Scaffold a custom AI agent harness via `metaharness new <name> --template <id> --host <id>`. Defaults to DRY-RUN (no writes) unless --confirm is passed. Refuses to write to the calling repo root or anywhere inside it. Honors ADR-150 architectural constraint + ruflo's "destructive-action confirmation" pattern.
|
|
4
|
+
argument-hint: "--name <id> --template <vertical:coding|minimal|…> [--host claude-code|codex|…] [--target /abs/path] [--confirm] [--format table|json]"
|
|
5
|
+
allowed-tools: Bash
|
|
6
|
+
---
|
|
7
|
+
|
|
8
|
+
The one write-capable skill in the plugin. Every other skill is
|
|
9
|
+
pure-read. This one calls `metaharness new`, which writes a new
|
|
10
|
+
directory tree.
|
|
11
|
+
|
|
12
|
+
## Safety (load-bearing)
|
|
13
|
+
|
|
14
|
+
1. **Dry-run by default.** Without `--confirm`, the script prints what
|
|
15
|
+
it would do and exits 0 without touching disk.
|
|
16
|
+
2. **Refuses project root.** If `--target` resolves to the current
|
|
17
|
+
working directory OR any path inside it, the script errors out with
|
|
18
|
+
exit 2. Target must be an absolute path OUTSIDE the calling repo
|
|
19
|
+
(default is a fresh `/tmp/ruflo-mint-<ts>-<name>/` dir).
|
|
20
|
+
3. **Refuses existing target.** Won't overwrite — must scaffold into a
|
|
21
|
+
non-existent dir.
|
|
22
|
+
4. **Subprocess + 60s timeout.** No library import, no in-process
|
|
23
|
+
execution. The mint stays sandboxed from ruflo's runtime.
|
|
24
|
+
|
|
25
|
+
## Algorithm
|
|
26
|
+
|
|
27
|
+
Implementation: [`scripts/mint.mjs`](../../scripts/mint.mjs).
|
|
28
|
+
|
|
29
|
+
1. Validate `--name`, `--template`. Default `--host` to `claude-code`.
|
|
30
|
+
2. Resolve `--target` (default: temp dir).
|
|
31
|
+
3. Run safety checks (no project-root writes; target must not exist).
|
|
32
|
+
4. Without `--confirm`: emit dry-run plan, exit 0.
|
|
33
|
+
5. With `--confirm`: shell `npx metaharness new <name> --template <id>
|
|
34
|
+
--host <id> --target <abs> --yes`.
|
|
35
|
+
|
|
36
|
+
## Templates
|
|
37
|
+
|
|
38
|
+
`minimal`, `vertical:coding`, `vertical:devops`, `vertical:support`,
|
|
39
|
+
`vertical:legal`, `vertical:research`, `vertical:trading`, `vertical:health`,
|
|
40
|
+
`vertical:education`, `vertical:sales`, `vertical:business`,
|
|
41
|
+
`vertical:crm`, `vertical:marketing`, `vertical:advertising`,
|
|
42
|
+
`vertical:ai`, `vertical:agentics`, `vertical:ruview`, `vertical:gaming`,
|
|
43
|
+
`vertical:repo-maintainer`, `vertical:exotic`.
|
|
44
|
+
|
|
45
|
+
## Hosts
|
|
46
|
+
|
|
47
|
+
`claude-code`, `codex`, `pi-dev`, `hermes`, `openclaw`, `rvm`,
|
|
48
|
+
`copilot`, `opencode`, `github-actions`.
|
|
49
|
+
|
|
50
|
+
## Example dry-run
|
|
51
|
+
|
|
52
|
+
```
|
|
53
|
+
$ node scripts/mint.mjs --name my-harness --template vertical:coding --host claude-code
|
|
54
|
+
# harness-mint (dry-run)
|
|
55
|
+
|
|
56
|
+
- action: metaharness new
|
|
57
|
+
- name: my-harness
|
|
58
|
+
- template: vertical:coding
|
|
59
|
+
- host: claude-code
|
|
60
|
+
- target: /tmp/ruflo-mint-1718560000-my-harness
|
|
61
|
+
- confirm: false
|
|
62
|
+
- willWrite: false
|
|
63
|
+
|
|
64
|
+
Re-run with `--confirm` to actually scaffold.
|
|
65
|
+
```
|
|
66
|
+
|
|
67
|
+
## Why dry-run by default
|
|
68
|
+
|
|
69
|
+
Ruflo's behavioral rules say "executing actions with care" — destructive
|
|
70
|
+
or repo-touching actions need confirmation. The dry-run output makes the
|
|
71
|
+
WHAT visible before the WHEN. A human sees `target`, decides, then
|
|
72
|
+
adds `--confirm` if happy.
|