jules-orchestrator-kit 0.38.0 β†’ 0.38.2

This diff represents the content of publicly available package versions that have been released to one of the supported registries. The information contained in this diff is provided for informational purposes only and reflects changes between package versions as they appear in their respective public registries.
Files changed (44) hide show
  1. package/.agent/rules/jules-protocol.md +3 -0
  2. package/JULES_RULES_TEMPLATE.md +3 -1
  3. package/README.md +163 -307
  4. package/bin/agentctl.mjs +177 -7
  5. package/bin/init.js +8 -8
  6. package/index.mjs +14 -2
  7. package/package.json +1 -1
  8. package/scripts/command-resolver.mjs +1 -2
  9. package/scripts/doc-sync-check.mjs +36 -1
  10. package/scripts/jules-create.mjs +1 -1
  11. package/scripts/jules-dispatch.mjs +1 -2
  12. package/scripts/jules-nightly.mjs +2 -2
  13. package/scripts/jules-patch.mjs +2 -2
  14. package/scripts/jules-queue-runner.mjs +1 -1
  15. package/scripts/jules-scan-todos.mjs +1 -1
  16. package/scripts/jules-self-audit.mjs +1 -2
  17. package/scripts/jules-status.mjs +1 -1
  18. package/scripts/release.mjs +55 -0
  19. package/scripts/run-tests.mjs +15 -1
  20. package/scripts/utils.mjs +14 -9
  21. package/src/asset-integrity.mjs +76 -0
  22. package/src/asset_integrity.mjs +1 -76
  23. package/src/config.mjs +42 -23
  24. package/src/dag-engine.mjs +44 -6
  25. package/src/engine.mjs +18 -11
  26. package/src/execution-envelope.mjs +111 -0
  27. package/src/execution_envelope.mjs +1 -111
  28. package/src/git.mjs +6 -3
  29. package/src/merge-blocks.mjs +178 -1
  30. package/src/ops/command-registry.mjs +25 -0
  31. package/src/ops/handover.mjs +411 -0
  32. package/src/ops/ide-scaffold.mjs +3 -3
  33. package/src/ops/next-step.mjs +6 -2
  34. package/src/prompt-guard.mjs +34 -3
  35. package/src/provider.mjs +191 -86
  36. package/src/role-resolver.mjs +31 -0
  37. package/src/rules-budget.mjs +179 -0
  38. package/src/rules_budget.mjs +1 -179
  39. package/src/security.mjs +98 -56
  40. package/src/state.mjs +11 -6
  41. package/src/task-optimizer.mjs +10 -2
  42. package/src/tui.mjs +1 -1
  43. package/src/webhook.mjs +2 -2
  44. package/src/wizard-task.mjs +4 -30
package/README.md CHANGED
@@ -1,8 +1,8 @@
1
1
  <div align="center">
2
2
 
3
- # πŸš€ jules-orchestrator-kit
3
+ # jules-orchestrator-kit
4
4
 
5
- ### Universal Autonomous AI Agent Orchestration Kernel for Google Jules
5
+ ### Task orchestration and automated verification harness for Google Jules
6
6
 
7
7
  <br/>
8
8
 
@@ -11,28 +11,24 @@
11
11
  [![License: MIT](https://img.shields.io/badge/License-MIT-yellow.svg)](https://opensource.org/licenses/MIT)
12
12
  [![Node.js Version](https://img.shields.io/badge/node-%3E%3D20.0.0-brightgreen.svg)](https://nodejs.org)
13
13
  [![Zero Dependencies](https://img.shields.io/badge/dependencies-0%20native-blue.svg)](https://nodejs.org)
14
+ [![Platform: Linux | macOS | Windows](https://img.shields.io/badge/platform-Linux%20%7C%20macOS%20%7C%20Windows-blueviolet.svg)](https://nodejs.org)
14
15
 
15
16
  <br/>
16
17
 
17
18
  <p align="center">
18
- <b>The zero-dependency safety gatekeeper and self-healing engineering kernel for autonomous coding agent swarms.</b>
19
- Transforms single-turn AI chat assistants into production-grade engineering swarms running 300+ daily sessions across any language or monorepo.
19
+ <b>Zero-dependency safety gatekeeper, scoped sandboxing, and automated verification for coding agents.</b><br/>
20
+ Runs deterministic test verification, secret scrubbing, and automated repair loops across any stack or monorepo before opening Pull Requests.
20
21
  </p>
21
22
 
22
23
  <br/>
23
24
 
24
25
  <p align="center">
25
- <a href="#what-is-kit">πŸ’‘&nbsp;What&nbsp;is&nbsp;Kit?</a> &nbsp;β€’&nbsp;
26
- <a href="#who-is-it-for">πŸ‘₯&nbsp;Who&nbsp;Is&nbsp;It&nbsp;For?</a> &nbsp;β€’&nbsp;
27
- <a href="#quickstart">⚑&nbsp;Quickstart</a> &nbsp;β€’&nbsp;
28
- <a href="#triage-guidelines">🎯&nbsp;Triage</a> &nbsp;β€’&nbsp;
29
- <a href="#matrix">πŸ“Š&nbsp;Matrix</a>
30
- <br/>
31
- <a href="#configuration">βš™οΈ&nbsp;Configuration</a> &nbsp;β€’&nbsp;
32
- <a href="#architecture">πŸ›οΈ&nbsp;Architecture</a> &nbsp;β€’&nbsp;
33
- <a href="#cli-docs">πŸ› οΈ&nbsp;CLI&nbsp;Docs</a> &nbsp;β€’&nbsp;
34
- <a href="#providers">πŸ”Œ&nbsp;Providers</a> &nbsp;β€’&nbsp;
35
- <a href="#roadmap">πŸ—ΊοΈ&nbsp;Roadmap</a>
26
+ <a href="#quickstart">Quickstart</a> &nbsp;β€’&nbsp;
27
+ <a href="#overview">Overview</a> &nbsp;β€’&nbsp;
28
+ <a href="#target-workflows">Target Workflows</a> &nbsp;β€’&nbsp;
29
+ <a href="#triage-guidelines">Triage</a> &nbsp;β€’&nbsp;
30
+ <a href="#cli-docs">CLI Docs</a> &nbsp;β€’&nbsp;
31
+ <a href="#deep-dives">Deep Dives</a>
36
32
  </p>
37
33
 
38
34
  </div>
@@ -49,32 +45,29 @@
49
45
 
50
46
  <br/>
51
47
 
52
- <a id="what-is-kit"></a>
53
- ## πŸ’‘ 2-Sentence Mental Model
54
-
55
- > [!TIP]
56
- > **Think of `jules-orchestrator-kit` as an automated Engineering Manager for AI coding agents.**
57
- > It hands out clear tasks, runs your tests in an isolated sandbox, fixes broken code automatically, and only opens a Pull Request when 100% of your tests pass.
58
-
59
- <br/>
60
-
61
- ---
48
+ <a id="quickstart"></a>
49
+ ## Quickstart
62
50
 
63
- <br/>
51
+ Get running in any repository in 3 commands (zero configuration required):
64
52
 
65
- <a id="who-is-it-for"></a>
66
- ## πŸ‘₯ Who Is It For?
53
+ ```bash
54
+ # 1. Initialize orchestrator in your project (auto-detects Python, Rust, Go, Node, PHP, etc.)
55
+ npx jules-orchestrator-kit init
67
56
 
68
- Whether you are trying your first AI coding session or running enterprise monorepo swarms, `jules-orchestrator-kit` scales with your workflow:
57
+ # 2. Author a scoped, verified task envelope with guardrails & secret scrubbing
58
+ npx jules-orchestrator-kit task create
69
59
 
70
- <br/>
60
+ # 3. Inspect repository health & diagnostic status
61
+ npx jules-orchestrator-kit doctor
62
+ ```
71
63
 
72
- | Role | Primary Value Proposition | Key Commands |
73
- | :--- | :--- | :--- |
74
- | **🌱 Beginners & Solo Developers** | Safely experiment with AI agents without risking broken code, leaked API keys, or ruined git history. | `agentctl init`<br/>`agentctl task create` |
75
- | **πŸ“¦ Single-Repo Maintainers** | Automate bug fixes, dependency updates, and PR reviews with automated OODA test verification. | `agentctl gate`<br/>`agentctl queue` |
76
- | **πŸ—οΈ Monorepo Engineering Teams** | Isolate subproject verification (`backend/`, `frontend/`, `cli/`) so agent edits never thrash global test suites. | `agentctl swarm`<br/>`agentctl lock` |
77
- | **πŸ›‘οΈ Platform & Security Engineers** | Enforce zero-trust security policies, pre-commit secret scrubbing, and strict 75 KB diff payload limits. | `agentctl doctor`<br/>`agentctl dashboard` |
64
+ > [!TIP]
65
+ > **Prefer a global CLI?**
66
+ > Install globally to access `agentctl` directly:
67
+ > ```bash
68
+ > npm install -g jules-orchestrator-kit
69
+ > agentctl init && agentctl task create && agentctl queue
70
+ > ```
78
71
 
79
72
  <br/>
80
73
 
@@ -82,53 +75,23 @@ Whether you are trying your first AI coding session or running enterprise monore
82
75
 
83
76
  <br/>
84
77
 
85
- ## 🎯 Why `jules-orchestrator-kit`?
86
-
87
- Autonomous coding agents can write software at 100Γ— human speedβ€”but unconstrained agents introduce silent regressions, leak API keys, hallucinate test assertions, and thrash shared monorepos.
88
-
89
- `jules-orchestrator-kit` provides the missing **Safety, Orchestration, and Verification Kernel** for high-reliability AI agent deployments:
90
-
91
- * **πŸ”₯ Warm Multi-Turn Session Resumption (`v0.31.0`):** `agentctl resume <sessionId> --response "<reply>"` streams an engineer's reply directly into an active Google Jules session via `POST /v1alpha/sessions/{id}:sendMessage`, preserving reasoning context instead of paying to rebuild it, with fail-soft cold-dispatch fallback on HTTP 400/404. This is the asynchronous HITL unblocking path; the automatic OODA repair loop currently opens a fresh session per attempt (see [architecture.md](docs/architecture.md#on-warm-session-resumption)).
92
-
93
- * **πŸ§ͺ Automated TDD Red-to-Green Harness (`agentctl test-gen`):** Scaffolds falsifiable unit tests from bug specs, verifies **RED** failure state, locks the test file in `scope.deny`, and tasks Jules with making it pass (**GREEN** state).
94
-
95
- * **πŸ›‘οΈ 1-Click Atomic Git Checkpoint & Rollback (`agentctl rollback`):** Snapshots working tree state, git diffs, and stashes before every session, enabling instant 1-command git restoration.
96
-
97
- * **🌐 Verification Sandbox & SSR Hydration Prober (`verify.server`):** Executes deterministic `setup`/`teardown` hooks for databases and boots dev servers to intercept React/Next.js SSR hydration panics before approving PRs.
98
-
99
- * **⚑ AST Blast-Radius Selective Testing:** Traverses file import dependency graphs to execute only affected downstream test suites, cutting monorepo test latency from minutes to milliseconds.
100
-
101
- * **πŸ”• Type III Silence Governor & Interruption Budgeting (`v0.35.0`):** `agentctl escalate` manages Slack and Discord webhook alerts with configurable digest modes and hourly interruption budgets, buffering non-critical notifications while guaranteeing zero-latency delivery for critical escalations.
102
-
103
- * **🩹 Automated Flaky Test Healing Swarm (`v0.35.0`):** `agentctl flaky heal` automatically consumes Wilson-quarantined tests (Exit Code 8) and dispatches specialized anti-flakiness repair tasks to eliminate race conditions, async timing leaks, and resource collisions without weakening test assertions.
78
+ <a id="overview"></a>
79
+ ## Overview
104
80
 
105
- * **πŸ”’ Zero Runtime Dependencies:** Built exclusively on Node.js 20+ built-ins (`node:fs`, `node:child_process`, `node:crypto`, `node:path`, `node:http`, `node:tty`, `node:test`). Zero third-party npm packages mean zero supply-chain CVE risk.
106
-
107
- * **πŸ›‘οΈ Fail-Closed Security Gatekeeper:** Unconditionally evaluates explicit Deny rules *before* Allow rules, matching against **canonicalised, case-folded paths** so `./`, `..`, mixed separators or a `.GitHub/` spelling cannot walk past a rule (the same repo is checked out on case-insensitive macOS and Windows filesystems). Redacts high-entropy secrets and PII from dry-runs and git diffs β€” **including credentials wrapped in base64**, so a key inside a Kubernetes `Secret` manifest is not invisible to a line-oriented scanner β€” blocks unsupported Node.js native module imports in Edge environments (Cloudflare Workers, Vercel Edge, Netlify Edge), and rejects PRs exceeding the 75 KB Diff Payload governor.
108
-
109
- * **πŸ”„ Autonomous OODA Self-Healing:** Captures test stderr/stdout, normalizes failure fingerprints, and feeds structured error contexts back into repair iterations (up to 3 automatic attempts) before human escalation.
110
-
111
- * **πŸ’» Native Interactive UX & Command Palette (`v0.30.0`):** Features a zero-dependency full-screen Terminal Engine (`capabilities`, `key-decoder`, `renderer`, `layout`, `widgets`), interactive diagnostic matrix (`agentctl doctor`), task queue/swarm managers (`agentctl queue`, `agentctl swarm`), and a searchable Command Palette.
112
-
113
- * **🌐 Universal Polyglot Spine:** Natively auto-detects 26+ tech stacks (PHP/Laravel/WordPress, .NET/C#, Python, Go, Rust, C/C++, Flutter/Swift, Node/Deno/Bun, Solidity/Foundry/Hardhat) and transparently wraps verification suites in Docker Compose or Devcontainer sandboxes.
114
-
115
- * **🧩 DAG-Ordered Task Queue & Specialist Agent Roles:** `agentctl queue --dag` resolves inter-task dependencies via Kahn's algorithm with cycle detection instead of linear FIFO order, and `--role <overseer|bolt|sentinel|janitor>` binds a task to a pre-defined specialist prompt persona resolved from `.agent/prompts/`.
116
-
117
- * **πŸ” Cryptographic Evidence Ledger (`agentctl evidence`):** Generates a SHA-256 manifest of changed files, test-file hashes, and tamper-detection locks for every verified task β€” a portable, offline-verifiable audit trail toward the roadmap's SOC2 compliance exporter.
118
-
119
- * **πŸ’Έ Dynamic Complexity & Cost Router (`router:` in `.agent/config.yml`, opt-in):** A zero-dependency, rule-based heuristic classifier (`src/router.mjs`) routes trivial tasks (typos, lint fixes, single-file lockfile bumps) to a cheap/fast provider β€” e.g. the built-in `gemini-flash` (Gemini CLI, headless mode) preset β€” while reserving your primary provider for complex, multi-file, or safety-sensitive work. Tasks touching `scope.deny`, `auth/**`, `migrations/**`, secrets, or using the `sentinel` role always force the primary provider regardless of score. Fully provider-agnostic: swap in any exec/HTTP provider spec for `router.fast`/`router.complex`, and override per-task with `--tier fast|complex`.
120
-
121
- * **πŸ“‚ Scoped Monorepo Boundary Resolver:** Statically maps changed files up directory ancestry to invoke isolated subshell test suites (`(cd backend && pytest) && (cd cli && cargo test)`), eliminating global test thrashing.
122
-
123
- * **πŸš€ Zero-Test Bootstrapping (`agentctl bootstrap`):** Synthesizes deterministic syntax-check and smoke-test verification oracles for untested legacy repositories so agents always operate against a falsifiable feedback loop.
124
-
125
- * **πŸ“ˆ Proven Scale & Reliability:** Empirically tested with **555 unit tests across 81 suites passing in < 10.0s**. An adversarial red-team suite (`test/adversarial-claims.test.mjs`) continuously attempts to falsify the safety guarantees documented above β€” every probe in it currently holds, with no open gaps β€” including cross-platform probes for the case-insensitive filesystems on macOS and Windows β€” and a documentation-sync gate (`scripts/doc-sync-check.mjs`) blocks any release whose docs have drifted from the code.
81
+ > **`jules-orchestrator-kit` serves as a safety gate and automated test runner for AI coding agents.**
82
+ > It drafts falsifiable task envelopes, executes verification commands in an isolated sandbox, automatically retries on test failures using captured diagnostics, and approves PRs only when 100% of tests pass cleanly.
126
83
 
127
84
  <br/>
128
85
 
129
- <p align="center">
130
- <img src="docs/assets/security-shield.svg" alt="Zero-Trust Security & Reliability Guarantees" width="100%" />
131
- </p>
86
+ <a id="target-workflows"></a>
87
+ ### Target Workflows
88
+
89
+ | Persona / Team | Primary Value | Everyday Commands |
90
+ | :--- | :--- | :--- |
91
+ | **Solo Developers** | Safely experiment with autonomous coding without risking broken branches, leaked API keys, or ruined git history. | `agentctl init`<br/>`agentctl task create` |
92
+ | **Repo Maintainers** | Automate bug fixes, dependency bumps, and PR reviews with self-healing test loops. | `agentctl gate`<br/>`agentctl queue` |
93
+ | **Monorepo Teams** | Isolate subproject verification (`backend/`, `frontend/`, `cli/`) so agent edits never thrash global test suites. | `agentctl swarm`<br/>`agentctl lock` |
94
+ | **Platform & Security** | Enforce fail-closed security policies, pre-commit secret scrubbing (including base64), and strict 75 KB diff limits. | `agentctl doctor`<br/>`agentctl dashboard` |
132
95
 
133
96
  <br/>
134
97
 
@@ -137,28 +100,22 @@ Autonomous coding agents can write software at 100Γ— human speedβ€”but unconstra
137
100
  <br/>
138
101
 
139
102
  <a id="triage-guidelines"></a>
140
- ## 🎯 Triage Guidelines: When to Use vs. When NOT to Use
141
-
142
- To ensure maximum merge success, dispatch tasks according to our deterministic triage boundaries:
103
+ ## Triage Guidelines: When to Dispatch Tasks
143
104
 
144
- <br/>
105
+ To maximize PR merge rates, dispatch tasks according to deterministic boundaries:
145
106
 
146
- ### 🟒 Ideal Tasks for Autonomous Swarms
107
+ ### Ideal Tasks (High Success Rate)
108
+ * **Scoped Bug Fixes & Code Changes:** Mechanically verifiable via unit tests (`pytest`, `npm test`, `cargo test`, `dotnet test`, `go test`).
109
+ * **Type & Linter Migrations:** Strict mode conversions, type annotations, and dead code elimination.
110
+ * **Dependency Bumps & CVE Patches:** Upgrading vulnerable lockfile dependencies with hermetic test validation.
111
+ * **Backend Refactoring:** Modularizing route controllers, API handlers, or database schemas.
112
+ * **Headless E2E / Playwright Tests:** UI changes verified by automated visual snapshots (`npx playwright test`).
147
113
 
148
- * βœ… **Scoped Code Changes & Bug Fixes:** Well-defined objectives mechanically verifiable via unit tests (`npm test`, `pytest`, `cargo test`, `dotnet test`).
149
- * βœ… **Type & Linter Migrations:** TypeScript strict mode fixes, PHP 8.3 type hint additions, or Python MyPy type annotation passes.
150
- * βœ… **Dependency Bumps & Security Audits:** Remediating CVEs in lockfiles (`package.json`, `Cargo.toml`, `composer.json`) with hermetic test verification.
151
- * βœ… **Refactoring Legacy Codebases:** Modularizing backend routes, API controllers, or database query layers.
152
- * βœ… **Visual & E2E Testing (via Playwright):** UI changes paired with automated headless Playwright snapshot tests (`npx playwright test`).
153
-
154
- <br/>
155
-
156
- ### πŸ”΄ When NOT to Use (Out of Scope)
157
-
158
- * ❌ **Unverifiable Visual UI Tweaks:** Pixel-perfect CSS/Tailwind adjustments lacking automated visual regression tests (agents cannot "see" raw browser output without Playwright).
159
- * ❌ **Closed Proprietary Platforms Without CLI:** Systems lacking local CLI tools or git repositories (e.g., Salesforce, Webflow, closed SAP backends).
160
- * ❌ **Unmocked Live Cloud Systems:** Code requiring live connections to 10+ external cloud APIs without local emulators or mocks.
161
- * ❌ **Protected Infrastructure Paths:** Direct edits to `.github/workflows/`, production deployment keys, or agent security gate rules (enforced fail-closed by `Agent Scope Guard`).
114
+ ### Out of Scope (Keep Human-in-the-Loop)
115
+ * **Unverifiable Visual UI Tweaks:** CSS/Tailwind adjustments without automated Playwright regression tests.
116
+ * **Closed Proprietary Platforms Without CLI:** Systems lacking local CLI or git integration (e.g. Salesforce GUI, Webflow).
117
+ * **Unmocked Live Cloud Systems:** Code requiring live connections to external cloud APIs without local mocks or emulators.
118
+ * **Protected Infrastructure Files:** Direct edits to `.github/workflows/`, deployment keys, or agent security gate rules (blocked fail-closed by `Agent Scope Guard`).
162
119
 
163
120
  <br/>
164
121
 
@@ -166,26 +123,15 @@ To ensure maximum merge success, dispatch tasks according to our deterministic t
166
123
 
167
124
  <br/>
168
125
 
169
- <a id="matrix"></a>
170
- ## πŸ“Š Feature Comparison Matrix
171
-
172
- | Dimension | Raw Agent Execution (No Orchestrator) | Standard CI/CD Pipelines | `jules-orchestrator-kit` (v0.31.0) |
173
- | :--- | :--- | :--- | :--- |
174
- | **Self-Healing Loop** | ❌ None (Crashes on test error) | ❌ None (Fails build; notifies human) | βœ… **Autonomous OODA Loop** (Max 3 repair turns with error fingerprinting) |
175
- | **Interactive UX Engine**| ❌ Raw unformatted CLI dumps | ❌ Non-interactive log outputs | βœ… **Native TUI Engine & Command Palette** (Zero-dependency alternate-screen TUI) |
176
- | **Scope Isolation** | ❌ None (Can modify CI files or lockfiles) | 🟑 Post-commit branch rules only | βœ… **Fail-Closed Scope Guard** (Deny-first evaluation; blocks protected paths) |
177
- | **Polyglot Stack Detection**| ❌ Manual prompt instructions | 🟑 Hardcoded YAML workflow steps | βœ… **Universal 26+ Stack Detector** (`src/config.mjs`) |
178
- | **Flaky Test Quarantine** | ❌ Fails session randomly | ❌ Breaks CI pipeline randomly | βœ… **Wilson-Score Statistical Quarantine** (Oscillation β‰₯ 0.40 quarantined automatically) |
179
- | **Monorepo Scoping** | ❌ Runs full global test suite | 🟑 Requires custom Nx/Turbo scripting | βœ… **Scoped Subshell Boundary Resolver** (`resolveWorkspaceBoundary`) |
180
- | **Zero-Test Bootstrapping**| ❌ Halts without verification oracle | ❌ Fails build if no tests exist | βœ… **Instant Oracle Synthesis** (`php -l`, `compileall`, `dotnet build`, `tsc`, `smoke`) |
181
- | **Secret Leak Prevention**| ❌ Prone to leaking tokens in diffs | 🟑 Post-push secret scanning alerts | βœ… **Pre-Dispatch & Pre-Commit Diff Scanner** (Blocks CVEs/keys before PR creation) |
182
- | **Dependency Footprint** | ❌ Requires heavy SDKs & parsers | 🟑 Many external actions & plugins | βœ… **0 Native Dependencies** (100% Node.js 20+ ESM built-ins) |
183
-
184
- <br/>
126
+ ## Core Capabilities
185
127
 
186
- <p align="center">
187
- <img src="docs/assets/tier-presets.svg" alt="Subscription Tier Allocation Matrix" width="100%" />
188
- </p>
128
+ * **Zero Runtime Dependencies:** Built exclusively on Node.js 20+ built-in modules (`node:fs`, `node:child_process`, `node:crypto`, `node:path`, `node:http`, `node:tty`, `node:test`).
129
+ * **Cross-Platform Parity:** Verified 100% green across Linux, macOS (Darwin), and Windows on Node 20, 22, and 24.
130
+ * **Autonomous Self-Healing Loop:** Captures test stderr/stdout, fingerprints error traces, and feeds structured context back into automated repair turns (up to 3 attempts) before human escalation.
131
+ * **Fail-Closed Security & Secret Redaction:** Evaluates explicit Deny rules before Allow rules against canonicalized, case-folded paths. Redacts high-entropy keys and base64-encoded credentials (such as Kubernetes `Secret` manifests).
132
+ * **Complexity & Cost Router:** Zero-dependency heuristic classifier (`src/router.mjs`) routing mechanical tasks to lightweight models while reserving primary models for complex refactors.
133
+ * **Terminal UI & Diagnostic Matrix (`agentctl doctor`):** Interactive terminal dashboard, task sidecar manager, and automated transactional self-repair.
134
+ * **Verified Test Suite:** Tested with **572 unit tests across 81 suites passing in < 10.0s**.
189
135
 
190
136
  <br/>
191
137
 
@@ -193,86 +139,44 @@ To ensure maximum merge success, dispatch tasks according to our deterministic t
193
139
 
194
140
  <br/>
195
141
 
196
- <a id="quickstart"></a>
197
- ## ⚑ Guided Quickstart (Audit β†’ Author β†’ Verify)
198
-
199
- Get started with a safe 3-step workflow across any repository:
142
+ <a id="cli-docs"></a>
143
+ ## CLI Command Reference (`agentctl`)
200
144
 
201
- <br/>
145
+ `agentctl` is the unified command-line interface for `jules-orchestrator-kit`, available via `npx jules-orchestrator-kit <command>` or `agentctl <command>`.
202
146
 
203
- ### Step 1: Security & Scope Gate Audit
204
- Audit your current working tree or branch for secret leaks, protected path violations, and verification readiness:
205
- ```bash
206
- # Run security, secret scanning, and scope gate audit without modifying files
207
- npx jules-orchestrator-kit gate --mode working-tree
208
- ```
147
+ | Command | Usage | Description | Exit Codes |
148
+ | :--- | :--- | :--- | :--- |
149
+ | `init` | `agentctl init [--interactive] [--tier pro]` | Interactive onboarding wizard & stack detector generating `.agent/config.yml`. | `0` (Created) |
150
+ | `task create` | `agentctl task create [--title <t>] [--prompt <p>] [--template <id>] [--role <name>] [--tier fast\|complex]` | Interactively authors & scopes falsifiable task envelopes with secret scrubbing, preflight gate checks, and DAG dependency wiring. | `0` (Queued), `1` (Secret/Unfalsifiable) |
151
+ | `task template` | `agentctl task template [<id>] [--list] [--json]` | Lists and synthesizes pre-calibrated web task envelopes (`web-cwv`, `web-wcag`, `web-seo`, `web-playwright`, `web-flaky-heal`, `web-i18n`, `web-ai-access`). | `0` (Listed/Synthesized) |
152
+ | `dispatch` | `agentctl dispatch [-p <prompt>] [-f <file>] [-r <role>] [-t <tier>] [--auto-pr] [--repoless] [--dry-run]` | Dispatches autonomous task to the active provider with payload limits and role prompt resolution. | `0` (Dispatched), `1` (Error) |
153
+ | `doctor` | `agentctl doctor [--json]` | Diagnostic DAG check runner & automated transactional self-repair engine. | `0` (Healthy), `1` (Failures) |
154
+ | `queue` | `agentctl queue [--dag] [--concurrency <n>] [--dry-run] [--json]` | Consumes and executes task envelopes in `.agent/jules-queue/` with Kahn's DAG dependency resolution. Non-task files (manifests, `README.md`) are skipped, and `--dry-run` previews without moving anything. | `0` (Complete) |
155
+ | `swarm` | `agentctl swarm [--json]` | Runs parallel multi-agent swarm across worker slots with PID liveness detection. | `0` (Complete) |
156
+ | `gate` / `audit`| `agentctl gate --mode working-tree [--json]` | Runs security, secret scanning, and verification gates against working tree or branch. | `0` (Approved), `3` (Scope), `5` (Diff >75K), `6` (Secret) |
157
+ | `rollback` | `agentctl rollback [sessionId \| --latest]` | Restores exact commit, uncommitted files, and cleans orphan task worktrees from pre-flight checkpoints. | `0` (Restored), `1` (Error) |
158
+ | `resume` | `agentctl resume <sessionId> --response "<reply>"` | Streams engineer response back into active Google Jules warm session context window. | `0` (Resumed), `1` (Error) |
159
+ | `test-gen` | `agentctl test-gen --title <t> --spec <s> [--run]` | Scaffolds falsifiable unit tests, verifies RED failure state, and locks test in `scope.deny`. | `0` (Scaffolded/Red) |
160
+ | `dashboard` | `agentctl dashboard [port]` | Starts zero-dependency local HTTP telemetry and audit visualizer dashboard. | `0` (Running) |
161
+ | `evidence` | `agentctl evidence <generate\|verify\|show>` | Generates, verifies, or prints SHA-256 cryptographic evidence manifests with test-tamper locking. | `0` (Verified), `1` (Tamper) |
162
+ | `flaky` | `agentctl flaky <status\|heal\|reset>` | Manages Wilson-quarantined tests (Exit Code 8) and dispatches automated anti-flakiness healing swarms. | `0` (Healed/Listed) |
163
+ | `mcp` | `agentctl mcp` | Starts stdio Model Context Protocol (MCP) server for Claude, Cursor, and Antigravity. | `0` / Stdio stream |
164
+ | `mcp init` | `agentctl mcp init [--target cursor\|vscode\|claude\|all]` | 1-click config scaffolding for Cursor (`.cursor/mcp.json`), VS Code tasks (`tasks.json`), and Claude Desktop. | `0` (Scaffolded) |
209
165
 
210
166
  <br/>
211
167
 
212
- ### Step 2: Author a Scoped Task
213
- Launch the interactive task authoring wizard or create a task via CLI:
214
- ```bash
215
- # Interactive authoring wizard with secret scrubbing & verification probes
216
- npx jules-orchestrator-kit task create
217
-
218
- # Or dispatch directly with explicit flags
219
- npx jules-orchestrator-kit task create --title "Fix authentication token expiration" \
220
- --prompt "Fix JWT expiration check in src/auth.mjs and verify with npm test."
221
- ```
168
+ ---
222
169
 
223
170
  <br/>
224
171
 
225
- ### Step 3: Local Verification & Queue Execution
226
- Inspect diagnostics and execute queued tasks:
227
- ```bash
228
- # Run interactive diagnostic matrix
229
- npx jules-orchestrator-kit doctor --interactive
230
-
231
- # Process pending task queue in isolated worktrees
232
- npx jules-orchestrator-kit queue --interactive
233
- ```
234
-
235
- <br/>
172
+ <a id="deep-dives"></a>
173
+ ## Deep Dives & Technical Reference
236
174
 
237
175
  <details>
238
- <summary><b>πŸ” View All 26+ Supported Ecosystems & Stack Triggers</b></summary>
176
+ <summary><b>Configuration Reference (<code>.agent/config.yml</code>)</b></summary>
239
177
 
240
178
  <br/>
241
179
 
242
- ```
243
- Ecosystems Natively Supported by src/config.mjs:
244
- β”œβ”€β”€ PHP / Laravel / WordPress (composer.json, phpunit.xml, pest.php, artisan, wp-cli.yml)
245
- β”œβ”€β”€ .NET / C# / F# (*.sln, *.csproj, *.fsproj, global.json)
246
- β”œβ”€β”€ Mobile / Dart / Flutter (pubspec.yaml)
247
- β”œβ”€β”€ Mobile / Swift / Xcode (Package.swift)
248
- β”œβ”€β”€ Mobile / React Native (app.json, react-native.config.js)
249
- β”œβ”€β”€ Systems / CMake (CMakeLists.txt)
250
- β”œβ”€β”€ Systems / Rust Cargo (Cargo.toml)
251
- β”œβ”€β”€ Systems / Go (go.mod)
252
- β”œβ”€β”€ Systems / Make (Makefile)
253
- β”œβ”€β”€ Web3 / Solidity Foundry (foundry.toml, remappings.txt) β€” offline-enforced forge test/build/fmt
254
- β”œβ”€β”€ Web3 / Solidity Hardhat (hardhat.config.js, hardhat.config.ts)
255
- β”œβ”€β”€ Python / FastAPI / Django (pyproject.toml, requirements.txt, setup.py)
256
- β”œβ”€β”€ Elixir / Phoenix (mix.exs)
257
- β”œβ”€β”€ Ruby / Rails (Gemfile)
258
- β”œβ”€β”€ Java / Maven (pom.xml)
259
- β”œβ”€β”€ Java / Gradle (build.gradle, build.gradle.kts)
260
- β”œβ”€β”€ JS / TS Workspaces (turbo.json, pnpm-workspace.yaml, nx.json)
261
- β”œβ”€β”€ JS / TS Runtimes (bunfig.toml, deno.json, package.json)
262
- └── Devcontainers & Docker Compose (.devcontainer/devcontainer.json, docker-compose.yml, Dockerfile)
263
- ```
264
-
265
- </details>
266
-
267
- <br/>
268
-
269
- ---
270
-
271
- <br/>
272
-
273
- <a id="configuration"></a>
274
- ## βš™οΈ Configuration Reference (`.agent/config.yml`)
275
-
276
180
  `jules-orchestrator-kit` auto-detects stack defaults, but allows explicit overrides through `.agent/config.yml`:
277
181
 
278
182
  ```yaml
@@ -283,7 +187,7 @@ provider: "jules" # Provider key ("jules" | "claude-code" | "codex" | "ge
283
187
  baseBranch: "main" # Default target base branch
284
188
  branchPrefix: "agent/" # Prefix for task branches
285
189
 
286
- # Verification commands (auto-detected by Stack Oracle if omitted)
190
+ # Verification commands (auto-detected by Stack Detector if omitted)
287
191
  verify:
288
192
  test: "npm test"
289
193
  build: "npm run build"
@@ -300,135 +204,86 @@ limits:
300
204
  diffKb: 75 # 75 KB Diff Payload Governor limit
301
205
  promptKb: 50 # Maximum prompt payload size
302
206
  dailyTasks: 300 # Task quota per rolling 24h window (not per calendar day)
303
- repairAttempts: 3 # Maximum OODA repair iterations
304
- concurrency: 15 # Worker slots; the ultra plan allows up to 60
207
+ repairAttempts: 3 # Maximum repair iterations
208
+ concurrency: 15 # Worker slots (free: 3, pro: 8, ultra: 15)
305
209
 
306
210
  # Dynamic Complexity & Cost Router β€” opt-in, disabled by default.
307
- # Provider-agnostic: "fast"/"complex" accept any provider key ("jules" |
308
- # "claude-code" | "codex" | "gemini-flash") or an inline custom provider spec.
309
211
  router:
310
212
  enabled: false
311
213
  fast: "gemini-flash" # Trivial/mechanical tasks (score <= threshold)
312
- complex: "jules" # Complex/multi-file/safety-sensitive tasks; defaults to `provider`
313
- threshold: 0 # Heuristic score above which a task escalates to `complex`
214
+ complex: "jules" # Complex/multi-file/safety-sensitive tasks
215
+ threshold: 0 # Heuristic score threshold for escalation
314
216
  ```
315
217
 
316
- <br/>
317
-
318
- ---
218
+ </details>
319
219
 
320
220
  <br/>
321
221
 
322
- <a id="architecture"></a>
323
- ## πŸ›οΈ System Architecture & Visual Diagrams
222
+ <details>
223
+ <summary><b>26+ Supported Languages, Frameworks & Stacks</b></summary>
324
224
 
325
225
  <br/>
326
226
 
327
- <p align="center">
328
- <img src="docs/assets/architecture-layers.svg" alt="Control Plane Architecture Layers" width="100%" />
329
- </p>
227
+ ```
228
+ Ecosystems Natively Detected & Verified by Stack Detector:
229
+ β”œβ”€β”€ Python / Django (pyproject.toml, requirements.txt, setup.py, manage.py)
230
+ β”œβ”€β”€ Systems / Rust Cargo (Cargo.toml)
231
+ β”œβ”€β”€ Systems / Go (go.mod)
232
+ β”œβ”€β”€ Systems / CMake & Make (CMakeLists.txt, Makefile)
233
+ β”œβ”€β”€ JS / TS Workspaces (turbo.json, pnpm-workspace.yaml, nx.json)
234
+ β”œβ”€β”€ JS / TS Runtimes (bunfig.toml, deno.json, package.json)
235
+ β”œβ”€β”€ PHP / Laravel / WordPress (composer.json, phpunit.xml, pest.php, artisan, wp-cli.yml)
236
+ β”œβ”€β”€ .NET / C# / F# (*.sln, *.csproj, *.fsproj, global.json)
237
+ β”œβ”€β”€ Mobile / Dart / Flutter (pubspec.yaml)
238
+ β”œβ”€β”€ Mobile / Swift / Xcode (Package.swift)
239
+ β”œβ”€β”€ Mobile / React Native (app.json, react-native.config.js)
240
+ β”œβ”€β”€ Web3 / Solidity Foundry (foundry.toml, remappings.txt) β€” offline-enforced
241
+ β”œβ”€β”€ Web3 / Solidity Hardhat (hardhat.config.js, hardhat.config.ts)
242
+ β”œβ”€β”€ Elixir / Phoenix (mix.exs)
243
+ β”œβ”€β”€ Ruby / Rails (Gemfile)
244
+ β”œβ”€β”€ Java / Maven & Gradle (pom.xml, build.gradle, build.gradle.kts)
245
+ └── Devcontainers & Docker Compose (.devcontainer/devcontainer.json, docker-compose.yml, Dockerfile)
246
+ ```
247
+
248
+ </details>
330
249
 
331
250
  <br/>
332
251
 
333
- ### 1. The Autonomous OODA Verification Loop
334
- Every task dispatched to `jules-orchestrator-kit` executes within an immutable, fail-closed verification loop:
252
+ <details>
253
+ <summary><b>System Architecture & Verification Diagrams</b></summary>
335
254
 
336
255
  <br/>
337
256
 
257
+ ### 1. Control Plane Architecture Layers
338
258
  <p align="center">
339
- <img src="docs/assets/ooda-loop-cycle.svg" alt="Self-Healing OODA Repair Loop & Thrash Breaker" width="100%" />
259
+ <img src="docs/assets/architecture-layers.svg" alt="Control Plane Architecture Layers" width="100%" />
340
260
  </p>
341
261
 
342
- <br/>
343
-
344
- ### 2. Polyglot Monorepo Scoped Execution Engine
345
- In monorepos containing multiple languages, `resolveWorkspaceBoundary(changedFiles)` traverses directory ancestry to isolate verification to affected subprojects:
346
-
347
- <br/>
348
-
262
+ ### 2. Autonomous Verification & Repair Loop
349
263
  <p align="center">
350
- <img src="docs/assets/monorepo-resolver.svg" alt="Polyglot Monorepo Scoped Boundary Resolver" width="100%" />
264
+ <img src="docs/assets/ooda-loop-cycle.svg" alt="Autonomous Verification & Repair Loop" width="100%" />
351
265
  </p>
352
266
 
353
- <br/>
354
-
355
- ### 3. Multi-Agent Parallel Swarm Topology
356
- Run concurrent agents across parallel worktree slots with AST/JSON 3-way merging:
357
-
358
- <br/>
359
-
267
+ ### 3. Polyglot Monorepo Scoped Boundary Resolver
360
268
  <p align="center">
361
- <img src="docs/assets/swarm-topology.svg" alt="Multi-Agent Parallel Swarm Topology" width="100%" />
269
+ <img src="docs/assets/monorepo-resolver.svg" alt="Polyglot Monorepo Scoped Boundary Resolver" width="100%" />
362
270
  </p>
363
271
 
364
- <br/>
365
-
366
- ### 4. Model Context Protocol (MCP) Integration
367
- Native stdio server exposing task dispatch, gate verification, and risk auditing to client tools (Antigravity, Claude, Cursor):
368
-
369
- <br/>
370
-
272
+ ### 4. Multi-Agent Parallel Swarm Topology
371
273
  <p align="center">
372
- <img src="docs/assets/mcp-integration.svg" alt="Model Context Protocol (MCP) Integration" width="100%" />
274
+ <img src="docs/assets/swarm-topology.svg" alt="Multi-Agent Parallel Swarm Topology" width="100%" />
373
275
  </p>
374
276
 
375
- <br/>
376
-
377
- ---
378
-
379
- <br/>
380
-
381
- <a id="cli-docs"></a>
382
- ## πŸ› οΈ CLI Command Reference (`agentctl`)
383
-
384
- `agentctl` is the unified command-line interface for `jules-orchestrator-kit`, available via `bin/agentctl.mjs` or `npx jules-orchestrator-kit <command>`.
385
-
386
- <br/>
387
-
388
- | Command | Usage | Description | Exit Codes |
389
- | :--- | :--- | :--- | :--- |
390
- | `init` | `agentctl init [--interactive] [--tier pro]` | Interactive onboarding wizard & stack oracle inspector generating `.agent/config.yml`. | `0` (Created) |
391
- | `task create` | `agentctl task create [--title <t>] [--prompt <p>] [--template <id>] [--role <name>] [--tier fast\|complex] [--depends-on <id,...>]` | Interactively authors & scopes falsifiable task envelopes with secret scrubbing, preflight gate checks, specialist role resolution, DAG dependency wiring, and an optional Cost Router tier override. | `0` (Queued), `1` (Unfalsifiable / Secret leak) |
392
- | `task template` | `agentctl task template [<id>] [--list] [--json]` | Lists and synthesizes specialized web task envelopes (`web-cwv`, `web-wcag`, `web-seo`, `web-playwright`, `web-flaky-heal`, `web-i18n`, `web-ai-access`). | `0` (Synthesized/Listed) |
393
- | `task optimize` | `agentctl task optimize "<prompt>" [--fix] [--web] [--json]` | Linter & optimizer injecting Google Labs 3-phase exploration budgets, critic steering, and web oracles. | `0` (Scored/Fixed) |
394
- | `test-gen` | `agentctl test-gen --title <t> --spec <s> [--run]` | Scaffolds falsifiable unit tests, verifies **RED** failure state, and locks test in `scope.deny`. | `0` (Scaffolded/Red) |
395
- | `rollback` | `agentctl rollback [sessionId \| --latest]` | Restores exact commit, uncommitted files, and cleans orphan task worktrees from pre-flight checkpoints. | `0` (Restored), `1` (Error) |
396
- | `resume` | `agentctl resume <sessionId> --response "<reply>"` | Streams engineer response back into active Google Jules warm session context window. | `0` (Resumed), `1` (Error) |
397
- | `dispatch` | `agentctl dispatch --title <t> --prompt <p> [--role <name>] [--tier fast\|complex]` | Dispatches a single task to an AI agent in an isolated worktree, optionally binding a specialist role prompt (`overseer`\|`bolt`\|`sentinel`\|`janitor`) and/or overriding the Cost Router tier. | `0` (Success), `1` (Arg error), `2` (429 Rate limit), `3` (Scope deny), `4` (OODA exhausted), `5` (Diff > 75KB), `6` (Secret leak) |
398
- | `doctor` | `agentctl doctor [--interactive] [--fix safe]` | Diagnostic DAG check runner & automated transactional repair planner. | `0` (Healthy) |
399
- | `queue` | `agentctl queue [--interactive] [--dag] [--concurrency <n>] [--json]` | Consumes, inspects, and executes task envelopes in `.agent/jules-queue/`; `--dag` resolves inter-task dependencies via Kahn's algorithm with cycle detection instead of linear FIFO order (supports `--json`). | `0` (Complete) |
400
- | `swarm` | `agentctl swarm [--interactive] [--json]` | Runs parallel multi-agent swarm across worker slots with process PID liveness detection (supports `--json`). | `0` (Complete) |
401
- | `scan` | `agentctl scan` | Scans codebase for TODO/FIXME annotations to seed task authoring. | `0` (Scanned) |
402
- | `review-repair`| `agentctl review-repair <pr-comments.json>`| Parses GitHub PR review comments and synthesizes actionable OODA repair tasks. | `0` (Parsed), `1` (Missing file) |
403
- | `dashboard` | `agentctl dashboard [port]` | Starts zero-dependency local HTTP telemetry and audit visualizer dashboard. | `0` (Running) |
404
- | `gate` / `audit`| `agentctl gate --mode working-tree [--json]` | Runs security, secret scanning, and verification gate against working tree or branch (supports `--json`). | `0` (Approved), `3` (Scope violation), `5` (Diff limit), `6` (Secret leak) |
405
- | `bootstrap` | `agentctl bootstrap [--force] [--json]` | Inspects an untested repository and synthesizes `.agent/config.yml` with a zero-test verification oracle (`php -l`, `compileall`, `dotnet build`, `tsc`, `smoke`). | `0` (Bootstrapped / Existing) |
406
- | `lock` | `agentctl lock <acquire\|release\|status>`| Manages VFS mutex locks for multi-agent non-overlapping file ownership. | `0` (Locked/Released), `1` (Conflict) |
407
- | `clean` | `agentctl clean` | Prunes stale git worktrees, lockfiles, and temporary ledgers. | `0` (Clean) |
408
- | `evidence` | `agentctl evidence <generate\|verify\|show> [--manifest <path>] [--json]` | Generates, verifies, or prints a SHA-256 cryptographic evidence manifest (changed-file hashes + test-file tamper lock) for audit trails. | `0` (Verified/Generated), `1` (Tamper detected / Verification failed) |
409
- | `escalate` | `agentctl escalate [<sessionId>] [--status] [--flush] [--clear]` | Dispatches or manages webhook escalation incidents across Slack and Discord with Type III Silence Governor and interruption budgeting. | `0` (Dispatched/Buffered) |
410
- | `flaky` | `agentctl flaky <status\|heal\|reset> [--dispatch] [--dry-run] [--role <name>]` | Manages Wilson-quarantined tests (Exit Code 8) and dispatches automated anti-flakiness healing swarm without assertion weakening. | `0` (Healed/Queued/Listed) |
411
- | `mcp` | `agentctl mcp` | Starts stdio Model Context Protocol (MCP) server for tool integration. | `0` / Stdio stream |
412
- | `mcp init` | `agentctl mcp init [--target cursor\|vscode\|claude\|all]` | 1-click scaffolding for Cursor (`.cursor/mcp.json`), VS Code tasks (`tasks.json`), and Claude Desktop. | `0` (Scaffolded) |
413
- | `version` | `agentctl version` | Outputs orchestrator kit semantic version. | `0` |
277
+ </details>
414
278
 
415
279
  <br/>
416
280
 
417
- ---
281
+ <details>
282
+ <summary><b>Multi-Provider Failover & Cost Router SDK</b></summary>
418
283
 
419
284
  <br/>
420
285
 
421
- <a id="providers"></a>
422
- ## πŸ”Œ Provider Integration & SDK Usage
423
-
424
- `jules-orchestrator-kit` supports standard AI agent platforms and programmatic failover routing:
425
-
426
- ### 1. Google Jules Native Integration
427
- Dispatch tasks using canonical task envelopes or the Google Jules REST v1alpha API.
428
-
429
- ### 2. Multi-Provider Failover SDK (`createFailoverProvider`)
430
- Programmatically configure ordered provider failover (e.g. falling back to secondary providers on HTTP 429 rate limits):
431
-
286
+ ### Multi-Provider Failover SDK (`createFailoverProvider`)
432
287
  ```javascript
433
288
  import { createFailoverProvider, loadConfig } from "jules-orchestrator-kit";
434
289
 
@@ -441,9 +296,7 @@ const result = await provider.dispatch(
441
296
  );
442
297
  ```
443
298
 
444
- ### 3. Dynamic Complexity & Cost Router (`resolveRoutedProvider`)
445
- Opt-in, config-driven routing between a cheap/fast provider and your primary provider β€” see [Configuration Reference](#configuration) for the `router:` block. Programmatic usage mirrors `createFailoverProvider`:
446
-
299
+ ### Cost Router SDK (`resolveRoutedProvider`)
447
300
  ```javascript
448
301
  import { resolveRoutedProvider, loadConfig } from "jules-orchestrator-kit";
449
302
 
@@ -455,41 +308,33 @@ const { provider, classification } = resolveRoutedProvider(
455
308
  console.log(classification.tier); // "fast" | "complex"
456
309
  ```
457
310
 
458
- Ships with a `gemini-flash` preset (Gemini CLI headless mode, `gemini-3.6-flash`) as a batteries-included fast tier, but any provider key or custom spec works for `router.fast`/`router.complex` β€” the router is provider-agnostic by design, not tied to any single vendor.
459
-
460
- ### 4. Model Context Protocol (MCP) Server
461
- Expose orchestrator gates and queue controls over stdio to client tools (Antigravity, Claude, Cursor):
462
- ```bash
463
- npx jules-orchestrator-kit mcp
464
- ```
311
+ </details>
465
312
 
466
313
  <br/>
467
314
 
468
- ---
315
+ <details>
316
+ <summary><b>Feature Roadmap & Shipped Milestones</b></summary>
469
317
 
470
318
  <br/>
471
319
 
472
- <a id="roadmap"></a>
473
- ## πŸ—ΊοΈ Feature Roadmap & Release History
474
-
475
- | Feature | Module / Command | Architectural Description | Target Release |
320
+ | Feature | Module / Command | Architectural Description | Status |
476
321
  | :--- | :--- | :--- | :---: |
477
- | **Dynamic Complexity & Cost Router** | `src/router.mjs`, `router:` in `.agent/config.yml` | Provider-agnostic, zero-dependency heuristic classifier routing trivial tasks to a fast/cheap provider (`gemini-flash` preset included) and complex/safety-sensitive tasks to the primary provider; opt-in, `--tier` override. | **v0.32.5** *(Shipped)* |
478
- | **DAG Task Queue, Specialist Roles & Evidence Ledger** | `src/dag-engine.mjs`, `src/evidence.mjs`, `agentctl evidence` | Kahn's-algorithm dependency-ordered queue execution (`queue --dag`), `--role` specialist prompt resolution, and SHA-256 cryptographic evidence manifests with test-tamper locking. | **v0.32.5** *(Shipped)* |
322
+ | **Queue Runner Fidelity** | `src/dag-engine.mjs`, `src/engine.mjs` | Queue selection is by task shape rather than file extension, so manifests and READMEs are skipped instead of dispatched, and `--dry-run` leaves the queue untouched. | **v0.38.2** *(Shipped)* |
323
+ | **Release Gate Enforcement & Wizard Smoke Test** | `.github/workflows/jules-audit.yml`, `scripts/release.mjs`, `test/wizard-smoke.test.mjs` | Doc-sync gate runs in CI rather than by hand, releases block on a green CI matrix for `HEAD`, per-test deadlines turn a hang into a failure, and the real `init` wizard is driven end to end over a fake TTY. | **v0.38.1** *(Shipped)* |
324
+ | **Multi-OS CI Matrix & TUI Hardening** | `scripts/run-tests.mjs`, `src/state.mjs`, `src/git.mjs` | Automated 9-job CI matrix across Linux, macOS, and Windows on Node 20/22/24 with raw-mode TUI resilience and native Windows command quoting. | **v0.38.0** *(Shipped)* |
325
+ | **Base64 Secret Detection & Budget Fix** | `src/security.mjs`, `src/budget.mjs` | Secret scanner decodes base64 before matching structured patterns (K8s secrets), and `budget reset` preserves confirmed provider sessions. | **v0.37.0** *(Shipped)* |
326
+ | **Universal AI Crawler Policy & llms.txt** | `src/web-templates.mjs` (`web-ai-access`) | Cross-surface consistency for crawler directives (`robots.txt`, meta tags, `X-Robots-Tag`) and `llms.txt` local route integrity. | **v0.36.0** *(Shipped)* |
327
+ | **Silence Governor & Flaky Test Swarm** | `src/webhook.mjs`, `src/flaky-ledger.mjs` | Notification alert throttling with interruption budgeting, and automated anti-flakiness swarm coordinator. | **v0.35.0** *(Shipped)* |
328
+ | **Rolling 24h Quota & Plan Concurrency** | `src/state.mjs`, `src/config.mjs` | Rolling 24-hour quota accounting matching vendor reset windows and true concurrency limits (3/15/60). | **v0.34.0** *(Shipped)* |
329
+ | **Cost Router & Guided First Run** | `src/router.mjs`, `src/ops/next-step.mjs` | Heuristic task classifier routing trivial tasks to fast models, and guided single-command first run workflow. | **v0.33.0** *(Shipped)* |
330
+ | **DAG Task Queue & Specialist Roles** | `src/dag-engine.mjs`, `src/evidence.mjs` | Kahn's-algorithm dependency queue execution (`queue --dag`), specialist role prompts (`overseer`, `bolt`, `sentinel`, `janitor`), and SHA-256 evidence manifests. | **v0.32.5** *(Shipped)* |
479
331
  | **Warm Session Resumption & PR Bundler** | `src/provider.mjs`, `src/engine.mjs` | Multi-turn warm session context streaming via `POST /v1alpha/sessions/{id}:sendMessage` & evidence PR descriptions. | **v0.31.0** *(Shipped)* |
480
332
  | **TDD Harness & Prompt Falsifiability Linter** | `agentctl test-gen`, `agentctl task optimize` | Automated RED-state test generator, `scope.deny` test locking, and prompt testability linter with fuzzy path resolution. | **v0.31.0** *(Shipped)* |
481
333
  | **Atomic Git Checkpoint & Rollback** | `agentctl rollback` (`src/ops/checkpoint.mjs`) | Pre-flight git HEAD/stash snapshotting, atomic rollback restoration, and 10-session pruning rotation. | **v0.31.0** *(Shipped)* |
482
- | **Verification Sandbox & SSR Hydration Prober** | `verify.server`, `verify.setup`/`teardown` | Isolated process group dev server probing, Next.js/React SSR panic detection, and deterministic DB hooks. | **v0.31.0** *(Shipped)* |
483
- | **AST Selective Testing & Escalation Bridge** | `src/dag-engine.mjs`, `agentctl escalate` | Downstream import test resolution, Slack/Discord webhook alerts, and async `agentctl resume` unblocking. | **v0.31.0** *(Shipped)* |
484
- | **IDE Native MCP Config Scaffolder** | `agentctl mcp init` (`src/ops/ide-scaffold.mjs`) | 1-click scaffolding for Cursor (`.cursor/mcp.json`), VS Code tasks (`tasks.json`), and Claude Desktop. | **v0.31.0** *(Shipped)* |
485
334
  | **Interactive UX Engine & TUI Engine** | `src/ux/` (`capabilities`, `key-decoder`, `renderer`, `layout`, `widgets`) | Zero-dependency terminal capabilities detector, sequence key decoder, virtual frame renderer, and widgets. | **v0.30.0** *(Shipped)* |
486
- | **Guided Diagnostics & Transactional Core** | `src/ops/` (`doctor-registry`, `doctor-planner`, `transaction`, `receipts`) | Diagnostic check DAG (`runDoctorChecks`), pure fix planner (`planDiagnosticFixes`), and transactional executor with rollback. | **v0.30.0** *(Shipped)* |
487
- | **Interactive Queue & Swarm Manager** | `src/ux/`, `src/ops/` (`queue-model`, `swarm-model`, `task-actions`, `swarm-actions`) | Task sidecar state machine, queue snapshot builder, PID liveness reconciler, task actions, and swarm actions. | **v0.30.0** *(Shipped)* |
488
- | **Command Registry & Command Palette** | `src/ops/command-registry.mjs`, `src/ux/palette.mjs` | Single-source command descriptor registry (`COMMAND_REGISTRY`), `--help` string formatter, fuzzy search filter, and command palette. | **v0.30.0** *(Shipped)* |
489
- | **Onboarding & Stack Oracle Wizard** | `agentctl init --interactive` (`src/wizard-init.mjs`) | Zero-dependency interactive CLI wizard auto-detecting verification oracles, quota tiers, and preset workflows. | **v0.29.0** *(Shipped)* |
490
- | **Guided Task Authoring Subsystem** | `agentctl task create` (`src/wizard-task.mjs`) | Guided task authoring with TODO candidate harvesting, Shannon entropy secret scrubbing, and guardrail footer synthesis. | **v0.29.0** *(Shipped)* |
491
- | **P0 Remediation & Safety Alignment** | Queue, Task Envelope & Secrets (`src/wizard-task.mjs`) | Canonical queue path alignment (`.agent/jules-queue/`), path traversal guards, atomic writes, multiline secret scans, and JSON headers. | **v0.29.1** *(Shipped)* |
492
- | **PR Review Auto-Remediation Loop** | `agentctl review-repair` (`src/review-repair.mjs`) | Ingests GitHub PR review comments (`CHANGES_REQUESTED`), extracts line/file context, and dispatches automated OODA repair turns. | **v0.27.0** *(Shipped)* |
335
+ | **PR Review Auto-Remediation Loop** | `agentctl review-repair` (`src/review-repair.mjs`) | Ingests GitHub PR review comments (`CHANGES_REQUESTED`), extracts line/file context, and dispatches automated repair turns. | **v0.27.0** *(Shipped)* |
336
+
337
+ </details>
493
338
 
494
339
  <br/>
495
340
 
@@ -497,9 +342,9 @@ npx jules-orchestrator-kit mcp
497
342
 
498
343
  <br/>
499
344
 
500
- ## πŸ“– Documentation & Architecture
345
+ ## πŸ“– Documentation & External References
501
346
 
502
- - [**System Architecture & Pipeline Overview**](./docs/architecture.md) β€” Comprehensive technical sequence diagram and control plane architecture.
347
+ - [**System Architecture & Pipeline Overview**](./docs/architecture.md) β€” Comprehensive technical sequence diagrams and control plane specifications.
503
348
  - [**Google Jules Official Documentation**](https://jules.google) β€” Official platform overview and API specifications for Google Jules.
504
349
  - [**Examples & Task Envelope Recipes**](./EXAMPLES.md) β€” Production YAML and Markdown task envelopes.
505
350
  - [**Changelog**](./CHANGELOG.md) β€” Full release history and migration guides.
@@ -510,6 +355,17 @@ npx jules-orchestrator-kit mcp
510
355
 
511
356
  <br/>
512
357
 
358
+ ## βš–οΈ Disclaimer
359
+
360
+ `jules-orchestrator-kit` is an independent, community-driven open-source project and is not affiliated with, endorsed by, or sponsored by Google, Google LLC, or Alphabet Inc. "Google", "Google Jules", and related marks are trademarks of Google LLC.
361
+
362
+ <br/>
363
+
364
+ ---
365
+
366
+ <br/>
367
+
513
368
  <div align="center">
514
- <p><b>jules-orchestrator-kit</b> β€’ Built with zero external dependencies for Google Jules and enterprise AI agent swarms.</p>
369
+ <p><b>jules-orchestrator-kit</b> β€’ Built with zero external dependencies for Google Jules and autonomous agent workflows.</p>
515
370
  </div>
371
+