jules-orchestrator-kit 0.38.0 β 0.38.2
This diff represents the content of publicly available package versions that have been released to one of the supported registries. The information contained in this diff is provided for informational purposes only and reflects changes between package versions as they appear in their respective public registries.
- package/.agent/rules/jules-protocol.md +3 -0
- package/JULES_RULES_TEMPLATE.md +3 -1
- package/README.md +163 -307
- package/bin/agentctl.mjs +177 -7
- package/bin/init.js +8 -8
- package/index.mjs +14 -2
- package/package.json +1 -1
- package/scripts/command-resolver.mjs +1 -2
- package/scripts/doc-sync-check.mjs +36 -1
- package/scripts/jules-create.mjs +1 -1
- package/scripts/jules-dispatch.mjs +1 -2
- package/scripts/jules-nightly.mjs +2 -2
- package/scripts/jules-patch.mjs +2 -2
- package/scripts/jules-queue-runner.mjs +1 -1
- package/scripts/jules-scan-todos.mjs +1 -1
- package/scripts/jules-self-audit.mjs +1 -2
- package/scripts/jules-status.mjs +1 -1
- package/scripts/release.mjs +55 -0
- package/scripts/run-tests.mjs +15 -1
- package/scripts/utils.mjs +14 -9
- package/src/asset-integrity.mjs +76 -0
- package/src/asset_integrity.mjs +1 -76
- package/src/config.mjs +42 -23
- package/src/dag-engine.mjs +44 -6
- package/src/engine.mjs +18 -11
- package/src/execution-envelope.mjs +111 -0
- package/src/execution_envelope.mjs +1 -111
- package/src/git.mjs +6 -3
- package/src/merge-blocks.mjs +178 -1
- package/src/ops/command-registry.mjs +25 -0
- package/src/ops/handover.mjs +411 -0
- package/src/ops/ide-scaffold.mjs +3 -3
- package/src/ops/next-step.mjs +6 -2
- package/src/prompt-guard.mjs +34 -3
- package/src/provider.mjs +191 -86
- package/src/role-resolver.mjs +31 -0
- package/src/rules-budget.mjs +179 -0
- package/src/rules_budget.mjs +1 -179
- package/src/security.mjs +98 -56
- package/src/state.mjs +11 -6
- package/src/task-optimizer.mjs +10 -2
- package/src/tui.mjs +1 -1
- package/src/webhook.mjs +2 -2
- package/src/wizard-task.mjs +4 -30
package/README.md
CHANGED
|
@@ -1,8 +1,8 @@
|
|
|
1
1
|
<div align="center">
|
|
2
2
|
|
|
3
|
-
#
|
|
3
|
+
# jules-orchestrator-kit
|
|
4
4
|
|
|
5
|
-
###
|
|
5
|
+
### Task orchestration and automated verification harness for Google Jules
|
|
6
6
|
|
|
7
7
|
<br/>
|
|
8
8
|
|
|
@@ -11,28 +11,24 @@
|
|
|
11
11
|
[](https://opensource.org/licenses/MIT)
|
|
12
12
|
[](https://nodejs.org)
|
|
13
13
|
[](https://nodejs.org)
|
|
14
|
+
[](https://nodejs.org)
|
|
14
15
|
|
|
15
16
|
<br/>
|
|
16
17
|
|
|
17
18
|
<p align="center">
|
|
18
|
-
<b>
|
|
19
|
-
|
|
19
|
+
<b>Zero-dependency safety gatekeeper, scoped sandboxing, and automated verification for coding agents.</b><br/>
|
|
20
|
+
Runs deterministic test verification, secret scrubbing, and automated repair loops across any stack or monorepo before opening Pull Requests.
|
|
20
21
|
</p>
|
|
21
22
|
|
|
22
23
|
<br/>
|
|
23
24
|
|
|
24
25
|
<p align="center">
|
|
25
|
-
<a href="#
|
|
26
|
-
<a href="#
|
|
27
|
-
<a href="#
|
|
28
|
-
<a href="#triage-guidelines"
|
|
29
|
-
<a href="#
|
|
30
|
-
<
|
|
31
|
-
<a href="#configuration">βοΈ Configuration</a> β’
|
|
32
|
-
<a href="#architecture">ποΈ Architecture</a> β’
|
|
33
|
-
<a href="#cli-docs">π οΈ CLI Docs</a> β’
|
|
34
|
-
<a href="#providers">π Providers</a> β’
|
|
35
|
-
<a href="#roadmap">πΊοΈ Roadmap</a>
|
|
26
|
+
<a href="#quickstart">Quickstart</a> β’
|
|
27
|
+
<a href="#overview">Overview</a> β’
|
|
28
|
+
<a href="#target-workflows">Target Workflows</a> β’
|
|
29
|
+
<a href="#triage-guidelines">Triage</a> β’
|
|
30
|
+
<a href="#cli-docs">CLI Docs</a> β’
|
|
31
|
+
<a href="#deep-dives">Deep Dives</a>
|
|
36
32
|
</p>
|
|
37
33
|
|
|
38
34
|
</div>
|
|
@@ -49,32 +45,29 @@
|
|
|
49
45
|
|
|
50
46
|
<br/>
|
|
51
47
|
|
|
52
|
-
<a id="
|
|
53
|
-
##
|
|
54
|
-
|
|
55
|
-
> [!TIP]
|
|
56
|
-
> **Think of `jules-orchestrator-kit` as an automated Engineering Manager for AI coding agents.**
|
|
57
|
-
> It hands out clear tasks, runs your tests in an isolated sandbox, fixes broken code automatically, and only opens a Pull Request when 100% of your tests pass.
|
|
58
|
-
|
|
59
|
-
<br/>
|
|
60
|
-
|
|
61
|
-
---
|
|
48
|
+
<a id="quickstart"></a>
|
|
49
|
+
## Quickstart
|
|
62
50
|
|
|
63
|
-
|
|
51
|
+
Get running in any repository in 3 commands (zero configuration required):
|
|
64
52
|
|
|
65
|
-
|
|
66
|
-
|
|
53
|
+
```bash
|
|
54
|
+
# 1. Initialize orchestrator in your project (auto-detects Python, Rust, Go, Node, PHP, etc.)
|
|
55
|
+
npx jules-orchestrator-kit init
|
|
67
56
|
|
|
68
|
-
|
|
57
|
+
# 2. Author a scoped, verified task envelope with guardrails & secret scrubbing
|
|
58
|
+
npx jules-orchestrator-kit task create
|
|
69
59
|
|
|
70
|
-
|
|
60
|
+
# 3. Inspect repository health & diagnostic status
|
|
61
|
+
npx jules-orchestrator-kit doctor
|
|
62
|
+
```
|
|
71
63
|
|
|
72
|
-
|
|
73
|
-
|
|
74
|
-
|
|
75
|
-
|
|
76
|
-
|
|
77
|
-
|
|
64
|
+
> [!TIP]
|
|
65
|
+
> **Prefer a global CLI?**
|
|
66
|
+
> Install globally to access `agentctl` directly:
|
|
67
|
+
> ```bash
|
|
68
|
+
> npm install -g jules-orchestrator-kit
|
|
69
|
+
> agentctl init && agentctl task create && agentctl queue
|
|
70
|
+
> ```
|
|
78
71
|
|
|
79
72
|
<br/>
|
|
80
73
|
|
|
@@ -82,53 +75,23 @@ Whether you are trying your first AI coding session or running enterprise monore
|
|
|
82
75
|
|
|
83
76
|
<br/>
|
|
84
77
|
|
|
85
|
-
|
|
86
|
-
|
|
87
|
-
Autonomous coding agents can write software at 100Γ human speedβbut unconstrained agents introduce silent regressions, leak API keys, hallucinate test assertions, and thrash shared monorepos.
|
|
88
|
-
|
|
89
|
-
`jules-orchestrator-kit` provides the missing **Safety, Orchestration, and Verification Kernel** for high-reliability AI agent deployments:
|
|
90
|
-
|
|
91
|
-
* **π₯ Warm Multi-Turn Session Resumption (`v0.31.0`):** `agentctl resume <sessionId> --response "<reply>"` streams an engineer's reply directly into an active Google Jules session via `POST /v1alpha/sessions/{id}:sendMessage`, preserving reasoning context instead of paying to rebuild it, with fail-soft cold-dispatch fallback on HTTP 400/404. This is the asynchronous HITL unblocking path; the automatic OODA repair loop currently opens a fresh session per attempt (see [architecture.md](docs/architecture.md#on-warm-session-resumption)).
|
|
92
|
-
|
|
93
|
-
* **π§ͺ Automated TDD Red-to-Green Harness (`agentctl test-gen`):** Scaffolds falsifiable unit tests from bug specs, verifies **RED** failure state, locks the test file in `scope.deny`, and tasks Jules with making it pass (**GREEN** state).
|
|
94
|
-
|
|
95
|
-
* **π‘οΈ 1-Click Atomic Git Checkpoint & Rollback (`agentctl rollback`):** Snapshots working tree state, git diffs, and stashes before every session, enabling instant 1-command git restoration.
|
|
96
|
-
|
|
97
|
-
* **π Verification Sandbox & SSR Hydration Prober (`verify.server`):** Executes deterministic `setup`/`teardown` hooks for databases and boots dev servers to intercept React/Next.js SSR hydration panics before approving PRs.
|
|
98
|
-
|
|
99
|
-
* **β‘ AST Blast-Radius Selective Testing:** Traverses file import dependency graphs to execute only affected downstream test suites, cutting monorepo test latency from minutes to milliseconds.
|
|
100
|
-
|
|
101
|
-
* **π Type III Silence Governor & Interruption Budgeting (`v0.35.0`):** `agentctl escalate` manages Slack and Discord webhook alerts with configurable digest modes and hourly interruption budgets, buffering non-critical notifications while guaranteeing zero-latency delivery for critical escalations.
|
|
102
|
-
|
|
103
|
-
* **π©Ή Automated Flaky Test Healing Swarm (`v0.35.0`):** `agentctl flaky heal` automatically consumes Wilson-quarantined tests (Exit Code 8) and dispatches specialized anti-flakiness repair tasks to eliminate race conditions, async timing leaks, and resource collisions without weakening test assertions.
|
|
78
|
+
<a id="overview"></a>
|
|
79
|
+
## Overview
|
|
104
80
|
|
|
105
|
-
|
|
106
|
-
|
|
107
|
-
* **π‘οΈ Fail-Closed Security Gatekeeper:** Unconditionally evaluates explicit Deny rules *before* Allow rules, matching against **canonicalised, case-folded paths** so `./`, `..`, mixed separators or a `.GitHub/` spelling cannot walk past a rule (the same repo is checked out on case-insensitive macOS and Windows filesystems). Redacts high-entropy secrets and PII from dry-runs and git diffs β **including credentials wrapped in base64**, so a key inside a Kubernetes `Secret` manifest is not invisible to a line-oriented scanner β blocks unsupported Node.js native module imports in Edge environments (Cloudflare Workers, Vercel Edge, Netlify Edge), and rejects PRs exceeding the 75 KB Diff Payload governor.
|
|
108
|
-
|
|
109
|
-
* **π Autonomous OODA Self-Healing:** Captures test stderr/stdout, normalizes failure fingerprints, and feeds structured error contexts back into repair iterations (up to 3 automatic attempts) before human escalation.
|
|
110
|
-
|
|
111
|
-
* **π» Native Interactive UX & Command Palette (`v0.30.0`):** Features a zero-dependency full-screen Terminal Engine (`capabilities`, `key-decoder`, `renderer`, `layout`, `widgets`), interactive diagnostic matrix (`agentctl doctor`), task queue/swarm managers (`agentctl queue`, `agentctl swarm`), and a searchable Command Palette.
|
|
112
|
-
|
|
113
|
-
* **π Universal Polyglot Spine:** Natively auto-detects 26+ tech stacks (PHP/Laravel/WordPress, .NET/C#, Python, Go, Rust, C/C++, Flutter/Swift, Node/Deno/Bun, Solidity/Foundry/Hardhat) and transparently wraps verification suites in Docker Compose or Devcontainer sandboxes.
|
|
114
|
-
|
|
115
|
-
* **π§© DAG-Ordered Task Queue & Specialist Agent Roles:** `agentctl queue --dag` resolves inter-task dependencies via Kahn's algorithm with cycle detection instead of linear FIFO order, and `--role <overseer|bolt|sentinel|janitor>` binds a task to a pre-defined specialist prompt persona resolved from `.agent/prompts/`.
|
|
116
|
-
|
|
117
|
-
* **π Cryptographic Evidence Ledger (`agentctl evidence`):** Generates a SHA-256 manifest of changed files, test-file hashes, and tamper-detection locks for every verified task β a portable, offline-verifiable audit trail toward the roadmap's SOC2 compliance exporter.
|
|
118
|
-
|
|
119
|
-
* **πΈ Dynamic Complexity & Cost Router (`router:` in `.agent/config.yml`, opt-in):** A zero-dependency, rule-based heuristic classifier (`src/router.mjs`) routes trivial tasks (typos, lint fixes, single-file lockfile bumps) to a cheap/fast provider β e.g. the built-in `gemini-flash` (Gemini CLI, headless mode) preset β while reserving your primary provider for complex, multi-file, or safety-sensitive work. Tasks touching `scope.deny`, `auth/**`, `migrations/**`, secrets, or using the `sentinel` role always force the primary provider regardless of score. Fully provider-agnostic: swap in any exec/HTTP provider spec for `router.fast`/`router.complex`, and override per-task with `--tier fast|complex`.
|
|
120
|
-
|
|
121
|
-
* **π Scoped Monorepo Boundary Resolver:** Statically maps changed files up directory ancestry to invoke isolated subshell test suites (`(cd backend && pytest) && (cd cli && cargo test)`), eliminating global test thrashing.
|
|
122
|
-
|
|
123
|
-
* **π Zero-Test Bootstrapping (`agentctl bootstrap`):** Synthesizes deterministic syntax-check and smoke-test verification oracles for untested legacy repositories so agents always operate against a falsifiable feedback loop.
|
|
124
|
-
|
|
125
|
-
* **π Proven Scale & Reliability:** Empirically tested with **555 unit tests across 81 suites passing in < 10.0s**. An adversarial red-team suite (`test/adversarial-claims.test.mjs`) continuously attempts to falsify the safety guarantees documented above β every probe in it currently holds, with no open gaps β including cross-platform probes for the case-insensitive filesystems on macOS and Windows β and a documentation-sync gate (`scripts/doc-sync-check.mjs`) blocks any release whose docs have drifted from the code.
|
|
81
|
+
> **`jules-orchestrator-kit` serves as a safety gate and automated test runner for AI coding agents.**
|
|
82
|
+
> It drafts falsifiable task envelopes, executes verification commands in an isolated sandbox, automatically retries on test failures using captured diagnostics, and approves PRs only when 100% of tests pass cleanly.
|
|
126
83
|
|
|
127
84
|
<br/>
|
|
128
85
|
|
|
129
|
-
<
|
|
130
|
-
|
|
131
|
-
|
|
86
|
+
<a id="target-workflows"></a>
|
|
87
|
+
### Target Workflows
|
|
88
|
+
|
|
89
|
+
| Persona / Team | Primary Value | Everyday Commands |
|
|
90
|
+
| :--- | :--- | :--- |
|
|
91
|
+
| **Solo Developers** | Safely experiment with autonomous coding without risking broken branches, leaked API keys, or ruined git history. | `agentctl init`<br/>`agentctl task create` |
|
|
92
|
+
| **Repo Maintainers** | Automate bug fixes, dependency bumps, and PR reviews with self-healing test loops. | `agentctl gate`<br/>`agentctl queue` |
|
|
93
|
+
| **Monorepo Teams** | Isolate subproject verification (`backend/`, `frontend/`, `cli/`) so agent edits never thrash global test suites. | `agentctl swarm`<br/>`agentctl lock` |
|
|
94
|
+
| **Platform & Security** | Enforce fail-closed security policies, pre-commit secret scrubbing (including base64), and strict 75 KB diff limits. | `agentctl doctor`<br/>`agentctl dashboard` |
|
|
132
95
|
|
|
133
96
|
<br/>
|
|
134
97
|
|
|
@@ -137,28 +100,22 @@ Autonomous coding agents can write software at 100Γ human speedβbut unconstra
|
|
|
137
100
|
<br/>
|
|
138
101
|
|
|
139
102
|
<a id="triage-guidelines"></a>
|
|
140
|
-
##
|
|
141
|
-
|
|
142
|
-
To ensure maximum merge success, dispatch tasks according to our deterministic triage boundaries:
|
|
103
|
+
## Triage Guidelines: When to Dispatch Tasks
|
|
143
104
|
|
|
144
|
-
|
|
105
|
+
To maximize PR merge rates, dispatch tasks according to deterministic boundaries:
|
|
145
106
|
|
|
146
|
-
###
|
|
107
|
+
### Ideal Tasks (High Success Rate)
|
|
108
|
+
* **Scoped Bug Fixes & Code Changes:** Mechanically verifiable via unit tests (`pytest`, `npm test`, `cargo test`, `dotnet test`, `go test`).
|
|
109
|
+
* **Type & Linter Migrations:** Strict mode conversions, type annotations, and dead code elimination.
|
|
110
|
+
* **Dependency Bumps & CVE Patches:** Upgrading vulnerable lockfile dependencies with hermetic test validation.
|
|
111
|
+
* **Backend Refactoring:** Modularizing route controllers, API handlers, or database schemas.
|
|
112
|
+
* **Headless E2E / Playwright Tests:** UI changes verified by automated visual snapshots (`npx playwright test`).
|
|
147
113
|
|
|
148
|
-
|
|
149
|
-
*
|
|
150
|
-
*
|
|
151
|
-
*
|
|
152
|
-
*
|
|
153
|
-
|
|
154
|
-
<br/>
|
|
155
|
-
|
|
156
|
-
### π΄ When NOT to Use (Out of Scope)
|
|
157
|
-
|
|
158
|
-
* β **Unverifiable Visual UI Tweaks:** Pixel-perfect CSS/Tailwind adjustments lacking automated visual regression tests (agents cannot "see" raw browser output without Playwright).
|
|
159
|
-
* β **Closed Proprietary Platforms Without CLI:** Systems lacking local CLI tools or git repositories (e.g., Salesforce, Webflow, closed SAP backends).
|
|
160
|
-
* β **Unmocked Live Cloud Systems:** Code requiring live connections to 10+ external cloud APIs without local emulators or mocks.
|
|
161
|
-
* β **Protected Infrastructure Paths:** Direct edits to `.github/workflows/`, production deployment keys, or agent security gate rules (enforced fail-closed by `Agent Scope Guard`).
|
|
114
|
+
### Out of Scope (Keep Human-in-the-Loop)
|
|
115
|
+
* **Unverifiable Visual UI Tweaks:** CSS/Tailwind adjustments without automated Playwright regression tests.
|
|
116
|
+
* **Closed Proprietary Platforms Without CLI:** Systems lacking local CLI or git integration (e.g. Salesforce GUI, Webflow).
|
|
117
|
+
* **Unmocked Live Cloud Systems:** Code requiring live connections to external cloud APIs without local mocks or emulators.
|
|
118
|
+
* **Protected Infrastructure Files:** Direct edits to `.github/workflows/`, deployment keys, or agent security gate rules (blocked fail-closed by `Agent Scope Guard`).
|
|
162
119
|
|
|
163
120
|
<br/>
|
|
164
121
|
|
|
@@ -166,26 +123,15 @@ To ensure maximum merge success, dispatch tasks according to our deterministic t
|
|
|
166
123
|
|
|
167
124
|
<br/>
|
|
168
125
|
|
|
169
|
-
|
|
170
|
-
## π Feature Comparison Matrix
|
|
171
|
-
|
|
172
|
-
| Dimension | Raw Agent Execution (No Orchestrator) | Standard CI/CD Pipelines | `jules-orchestrator-kit` (v0.31.0) |
|
|
173
|
-
| :--- | :--- | :--- | :--- |
|
|
174
|
-
| **Self-Healing Loop** | β None (Crashes on test error) | β None (Fails build; notifies human) | β
**Autonomous OODA Loop** (Max 3 repair turns with error fingerprinting) |
|
|
175
|
-
| **Interactive UX Engine**| β Raw unformatted CLI dumps | β Non-interactive log outputs | β
**Native TUI Engine & Command Palette** (Zero-dependency alternate-screen TUI) |
|
|
176
|
-
| **Scope Isolation** | β None (Can modify CI files or lockfiles) | π‘ Post-commit branch rules only | β
**Fail-Closed Scope Guard** (Deny-first evaluation; blocks protected paths) |
|
|
177
|
-
| **Polyglot Stack Detection**| β Manual prompt instructions | π‘ Hardcoded YAML workflow steps | β
**Universal 26+ Stack Detector** (`src/config.mjs`) |
|
|
178
|
-
| **Flaky Test Quarantine** | β Fails session randomly | β Breaks CI pipeline randomly | β
**Wilson-Score Statistical Quarantine** (Oscillation β₯ 0.40 quarantined automatically) |
|
|
179
|
-
| **Monorepo Scoping** | β Runs full global test suite | π‘ Requires custom Nx/Turbo scripting | β
**Scoped Subshell Boundary Resolver** (`resolveWorkspaceBoundary`) |
|
|
180
|
-
| **Zero-Test Bootstrapping**| β Halts without verification oracle | β Fails build if no tests exist | β
**Instant Oracle Synthesis** (`php -l`, `compileall`, `dotnet build`, `tsc`, `smoke`) |
|
|
181
|
-
| **Secret Leak Prevention**| β Prone to leaking tokens in diffs | π‘ Post-push secret scanning alerts | β
**Pre-Dispatch & Pre-Commit Diff Scanner** (Blocks CVEs/keys before PR creation) |
|
|
182
|
-
| **Dependency Footprint** | β Requires heavy SDKs & parsers | π‘ Many external actions & plugins | β
**0 Native Dependencies** (100% Node.js 20+ ESM built-ins) |
|
|
183
|
-
|
|
184
|
-
<br/>
|
|
126
|
+
## Core Capabilities
|
|
185
127
|
|
|
186
|
-
|
|
187
|
-
|
|
188
|
-
|
|
128
|
+
* **Zero Runtime Dependencies:** Built exclusively on Node.js 20+ built-in modules (`node:fs`, `node:child_process`, `node:crypto`, `node:path`, `node:http`, `node:tty`, `node:test`).
|
|
129
|
+
* **Cross-Platform Parity:** Verified 100% green across Linux, macOS (Darwin), and Windows on Node 20, 22, and 24.
|
|
130
|
+
* **Autonomous Self-Healing Loop:** Captures test stderr/stdout, fingerprints error traces, and feeds structured context back into automated repair turns (up to 3 attempts) before human escalation.
|
|
131
|
+
* **Fail-Closed Security & Secret Redaction:** Evaluates explicit Deny rules before Allow rules against canonicalized, case-folded paths. Redacts high-entropy keys and base64-encoded credentials (such as Kubernetes `Secret` manifests).
|
|
132
|
+
* **Complexity & Cost Router:** Zero-dependency heuristic classifier (`src/router.mjs`) routing mechanical tasks to lightweight models while reserving primary models for complex refactors.
|
|
133
|
+
* **Terminal UI & Diagnostic Matrix (`agentctl doctor`):** Interactive terminal dashboard, task sidecar manager, and automated transactional self-repair.
|
|
134
|
+
* **Verified Test Suite:** Tested with **572 unit tests across 81 suites passing in < 10.0s**.
|
|
189
135
|
|
|
190
136
|
<br/>
|
|
191
137
|
|
|
@@ -193,86 +139,44 @@ To ensure maximum merge success, dispatch tasks according to our deterministic t
|
|
|
193
139
|
|
|
194
140
|
<br/>
|
|
195
141
|
|
|
196
|
-
<a id="
|
|
197
|
-
##
|
|
198
|
-
|
|
199
|
-
Get started with a safe 3-step workflow across any repository:
|
|
142
|
+
<a id="cli-docs"></a>
|
|
143
|
+
## CLI Command Reference (`agentctl`)
|
|
200
144
|
|
|
201
|
-
<
|
|
145
|
+
`agentctl` is the unified command-line interface for `jules-orchestrator-kit`, available via `npx jules-orchestrator-kit <command>` or `agentctl <command>`.
|
|
202
146
|
|
|
203
|
-
|
|
204
|
-
|
|
205
|
-
|
|
206
|
-
|
|
207
|
-
|
|
208
|
-
|
|
147
|
+
| Command | Usage | Description | Exit Codes |
|
|
148
|
+
| :--- | :--- | :--- | :--- |
|
|
149
|
+
| `init` | `agentctl init [--interactive] [--tier pro]` | Interactive onboarding wizard & stack detector generating `.agent/config.yml`. | `0` (Created) |
|
|
150
|
+
| `task create` | `agentctl task create [--title <t>] [--prompt <p>] [--template <id>] [--role <name>] [--tier fast\|complex]` | Interactively authors & scopes falsifiable task envelopes with secret scrubbing, preflight gate checks, and DAG dependency wiring. | `0` (Queued), `1` (Secret/Unfalsifiable) |
|
|
151
|
+
| `task template` | `agentctl task template [<id>] [--list] [--json]` | Lists and synthesizes pre-calibrated web task envelopes (`web-cwv`, `web-wcag`, `web-seo`, `web-playwright`, `web-flaky-heal`, `web-i18n`, `web-ai-access`). | `0` (Listed/Synthesized) |
|
|
152
|
+
| `dispatch` | `agentctl dispatch [-p <prompt>] [-f <file>] [-r <role>] [-t <tier>] [--auto-pr] [--repoless] [--dry-run]` | Dispatches autonomous task to the active provider with payload limits and role prompt resolution. | `0` (Dispatched), `1` (Error) |
|
|
153
|
+
| `doctor` | `agentctl doctor [--json]` | Diagnostic DAG check runner & automated transactional self-repair engine. | `0` (Healthy), `1` (Failures) |
|
|
154
|
+
| `queue` | `agentctl queue [--dag] [--concurrency <n>] [--dry-run] [--json]` | Consumes and executes task envelopes in `.agent/jules-queue/` with Kahn's DAG dependency resolution. Non-task files (manifests, `README.md`) are skipped, and `--dry-run` previews without moving anything. | `0` (Complete) |
|
|
155
|
+
| `swarm` | `agentctl swarm [--json]` | Runs parallel multi-agent swarm across worker slots with PID liveness detection. | `0` (Complete) |
|
|
156
|
+
| `gate` / `audit`| `agentctl gate --mode working-tree [--json]` | Runs security, secret scanning, and verification gates against working tree or branch. | `0` (Approved), `3` (Scope), `5` (Diff >75K), `6` (Secret) |
|
|
157
|
+
| `rollback` | `agentctl rollback [sessionId \| --latest]` | Restores exact commit, uncommitted files, and cleans orphan task worktrees from pre-flight checkpoints. | `0` (Restored), `1` (Error) |
|
|
158
|
+
| `resume` | `agentctl resume <sessionId> --response "<reply>"` | Streams engineer response back into active Google Jules warm session context window. | `0` (Resumed), `1` (Error) |
|
|
159
|
+
| `test-gen` | `agentctl test-gen --title <t> --spec <s> [--run]` | Scaffolds falsifiable unit tests, verifies RED failure state, and locks test in `scope.deny`. | `0` (Scaffolded/Red) |
|
|
160
|
+
| `dashboard` | `agentctl dashboard [port]` | Starts zero-dependency local HTTP telemetry and audit visualizer dashboard. | `0` (Running) |
|
|
161
|
+
| `evidence` | `agentctl evidence <generate\|verify\|show>` | Generates, verifies, or prints SHA-256 cryptographic evidence manifests with test-tamper locking. | `0` (Verified), `1` (Tamper) |
|
|
162
|
+
| `flaky` | `agentctl flaky <status\|heal\|reset>` | Manages Wilson-quarantined tests (Exit Code 8) and dispatches automated anti-flakiness healing swarms. | `0` (Healed/Listed) |
|
|
163
|
+
| `mcp` | `agentctl mcp` | Starts stdio Model Context Protocol (MCP) server for Claude, Cursor, and Antigravity. | `0` / Stdio stream |
|
|
164
|
+
| `mcp init` | `agentctl mcp init [--target cursor\|vscode\|claude\|all]` | 1-click config scaffolding for Cursor (`.cursor/mcp.json`), VS Code tasks (`tasks.json`), and Claude Desktop. | `0` (Scaffolded) |
|
|
209
165
|
|
|
210
166
|
<br/>
|
|
211
167
|
|
|
212
|
-
|
|
213
|
-
Launch the interactive task authoring wizard or create a task via CLI:
|
|
214
|
-
```bash
|
|
215
|
-
# Interactive authoring wizard with secret scrubbing & verification probes
|
|
216
|
-
npx jules-orchestrator-kit task create
|
|
217
|
-
|
|
218
|
-
# Or dispatch directly with explicit flags
|
|
219
|
-
npx jules-orchestrator-kit task create --title "Fix authentication token expiration" \
|
|
220
|
-
--prompt "Fix JWT expiration check in src/auth.mjs and verify with npm test."
|
|
221
|
-
```
|
|
168
|
+
---
|
|
222
169
|
|
|
223
170
|
<br/>
|
|
224
171
|
|
|
225
|
-
|
|
226
|
-
|
|
227
|
-
```bash
|
|
228
|
-
# Run interactive diagnostic matrix
|
|
229
|
-
npx jules-orchestrator-kit doctor --interactive
|
|
230
|
-
|
|
231
|
-
# Process pending task queue in isolated worktrees
|
|
232
|
-
npx jules-orchestrator-kit queue --interactive
|
|
233
|
-
```
|
|
234
|
-
|
|
235
|
-
<br/>
|
|
172
|
+
<a id="deep-dives"></a>
|
|
173
|
+
## Deep Dives & Technical Reference
|
|
236
174
|
|
|
237
175
|
<details>
|
|
238
|
-
<summary><b
|
|
176
|
+
<summary><b>Configuration Reference (<code>.agent/config.yml</code>)</b></summary>
|
|
239
177
|
|
|
240
178
|
<br/>
|
|
241
179
|
|
|
242
|
-
```
|
|
243
|
-
Ecosystems Natively Supported by src/config.mjs:
|
|
244
|
-
βββ PHP / Laravel / WordPress (composer.json, phpunit.xml, pest.php, artisan, wp-cli.yml)
|
|
245
|
-
βββ .NET / C# / F# (*.sln, *.csproj, *.fsproj, global.json)
|
|
246
|
-
βββ Mobile / Dart / Flutter (pubspec.yaml)
|
|
247
|
-
βββ Mobile / Swift / Xcode (Package.swift)
|
|
248
|
-
βββ Mobile / React Native (app.json, react-native.config.js)
|
|
249
|
-
βββ Systems / CMake (CMakeLists.txt)
|
|
250
|
-
βββ Systems / Rust Cargo (Cargo.toml)
|
|
251
|
-
βββ Systems / Go (go.mod)
|
|
252
|
-
βββ Systems / Make (Makefile)
|
|
253
|
-
βββ Web3 / Solidity Foundry (foundry.toml, remappings.txt) β offline-enforced forge test/build/fmt
|
|
254
|
-
βββ Web3 / Solidity Hardhat (hardhat.config.js, hardhat.config.ts)
|
|
255
|
-
βββ Python / FastAPI / Django (pyproject.toml, requirements.txt, setup.py)
|
|
256
|
-
βββ Elixir / Phoenix (mix.exs)
|
|
257
|
-
βββ Ruby / Rails (Gemfile)
|
|
258
|
-
βββ Java / Maven (pom.xml)
|
|
259
|
-
βββ Java / Gradle (build.gradle, build.gradle.kts)
|
|
260
|
-
βββ JS / TS Workspaces (turbo.json, pnpm-workspace.yaml, nx.json)
|
|
261
|
-
βββ JS / TS Runtimes (bunfig.toml, deno.json, package.json)
|
|
262
|
-
βββ Devcontainers & Docker Compose (.devcontainer/devcontainer.json, docker-compose.yml, Dockerfile)
|
|
263
|
-
```
|
|
264
|
-
|
|
265
|
-
</details>
|
|
266
|
-
|
|
267
|
-
<br/>
|
|
268
|
-
|
|
269
|
-
---
|
|
270
|
-
|
|
271
|
-
<br/>
|
|
272
|
-
|
|
273
|
-
<a id="configuration"></a>
|
|
274
|
-
## βοΈ Configuration Reference (`.agent/config.yml`)
|
|
275
|
-
|
|
276
180
|
`jules-orchestrator-kit` auto-detects stack defaults, but allows explicit overrides through `.agent/config.yml`:
|
|
277
181
|
|
|
278
182
|
```yaml
|
|
@@ -283,7 +187,7 @@ provider: "jules" # Provider key ("jules" | "claude-code" | "codex" | "ge
|
|
|
283
187
|
baseBranch: "main" # Default target base branch
|
|
284
188
|
branchPrefix: "agent/" # Prefix for task branches
|
|
285
189
|
|
|
286
|
-
# Verification commands (auto-detected by Stack
|
|
190
|
+
# Verification commands (auto-detected by Stack Detector if omitted)
|
|
287
191
|
verify:
|
|
288
192
|
test: "npm test"
|
|
289
193
|
build: "npm run build"
|
|
@@ -300,135 +204,86 @@ limits:
|
|
|
300
204
|
diffKb: 75 # 75 KB Diff Payload Governor limit
|
|
301
205
|
promptKb: 50 # Maximum prompt payload size
|
|
302
206
|
dailyTasks: 300 # Task quota per rolling 24h window (not per calendar day)
|
|
303
|
-
repairAttempts: 3 # Maximum
|
|
304
|
-
concurrency: 15 # Worker slots
|
|
207
|
+
repairAttempts: 3 # Maximum repair iterations
|
|
208
|
+
concurrency: 15 # Worker slots (free: 3, pro: 8, ultra: 15)
|
|
305
209
|
|
|
306
210
|
# Dynamic Complexity & Cost Router β opt-in, disabled by default.
|
|
307
|
-
# Provider-agnostic: "fast"/"complex" accept any provider key ("jules" |
|
|
308
|
-
# "claude-code" | "codex" | "gemini-flash") or an inline custom provider spec.
|
|
309
211
|
router:
|
|
310
212
|
enabled: false
|
|
311
213
|
fast: "gemini-flash" # Trivial/mechanical tasks (score <= threshold)
|
|
312
|
-
complex: "jules" # Complex/multi-file/safety-sensitive tasks
|
|
313
|
-
threshold: 0 # Heuristic score
|
|
214
|
+
complex: "jules" # Complex/multi-file/safety-sensitive tasks
|
|
215
|
+
threshold: 0 # Heuristic score threshold for escalation
|
|
314
216
|
```
|
|
315
217
|
|
|
316
|
-
|
|
317
|
-
|
|
318
|
-
---
|
|
218
|
+
</details>
|
|
319
219
|
|
|
320
220
|
<br/>
|
|
321
221
|
|
|
322
|
-
<
|
|
323
|
-
|
|
222
|
+
<details>
|
|
223
|
+
<summary><b>26+ Supported Languages, Frameworks & Stacks</b></summary>
|
|
324
224
|
|
|
325
225
|
<br/>
|
|
326
226
|
|
|
327
|
-
|
|
328
|
-
|
|
329
|
-
|
|
227
|
+
```
|
|
228
|
+
Ecosystems Natively Detected & Verified by Stack Detector:
|
|
229
|
+
βββ Python / Django (pyproject.toml, requirements.txt, setup.py, manage.py)
|
|
230
|
+
βββ Systems / Rust Cargo (Cargo.toml)
|
|
231
|
+
βββ Systems / Go (go.mod)
|
|
232
|
+
βββ Systems / CMake & Make (CMakeLists.txt, Makefile)
|
|
233
|
+
βββ JS / TS Workspaces (turbo.json, pnpm-workspace.yaml, nx.json)
|
|
234
|
+
βββ JS / TS Runtimes (bunfig.toml, deno.json, package.json)
|
|
235
|
+
βββ PHP / Laravel / WordPress (composer.json, phpunit.xml, pest.php, artisan, wp-cli.yml)
|
|
236
|
+
βββ .NET / C# / F# (*.sln, *.csproj, *.fsproj, global.json)
|
|
237
|
+
βββ Mobile / Dart / Flutter (pubspec.yaml)
|
|
238
|
+
βββ Mobile / Swift / Xcode (Package.swift)
|
|
239
|
+
βββ Mobile / React Native (app.json, react-native.config.js)
|
|
240
|
+
βββ Web3 / Solidity Foundry (foundry.toml, remappings.txt) β offline-enforced
|
|
241
|
+
βββ Web3 / Solidity Hardhat (hardhat.config.js, hardhat.config.ts)
|
|
242
|
+
βββ Elixir / Phoenix (mix.exs)
|
|
243
|
+
βββ Ruby / Rails (Gemfile)
|
|
244
|
+
βββ Java / Maven & Gradle (pom.xml, build.gradle, build.gradle.kts)
|
|
245
|
+
βββ Devcontainers & Docker Compose (.devcontainer/devcontainer.json, docker-compose.yml, Dockerfile)
|
|
246
|
+
```
|
|
247
|
+
|
|
248
|
+
</details>
|
|
330
249
|
|
|
331
250
|
<br/>
|
|
332
251
|
|
|
333
|
-
|
|
334
|
-
|
|
252
|
+
<details>
|
|
253
|
+
<summary><b>System Architecture & Verification Diagrams</b></summary>
|
|
335
254
|
|
|
336
255
|
<br/>
|
|
337
256
|
|
|
257
|
+
### 1. Control Plane Architecture Layers
|
|
338
258
|
<p align="center">
|
|
339
|
-
<img src="docs/assets/
|
|
259
|
+
<img src="docs/assets/architecture-layers.svg" alt="Control Plane Architecture Layers" width="100%" />
|
|
340
260
|
</p>
|
|
341
261
|
|
|
342
|
-
|
|
343
|
-
|
|
344
|
-
### 2. Polyglot Monorepo Scoped Execution Engine
|
|
345
|
-
In monorepos containing multiple languages, `resolveWorkspaceBoundary(changedFiles)` traverses directory ancestry to isolate verification to affected subprojects:
|
|
346
|
-
|
|
347
|
-
<br/>
|
|
348
|
-
|
|
262
|
+
### 2. Autonomous Verification & Repair Loop
|
|
349
263
|
<p align="center">
|
|
350
|
-
<img src="docs/assets/
|
|
264
|
+
<img src="docs/assets/ooda-loop-cycle.svg" alt="Autonomous Verification & Repair Loop" width="100%" />
|
|
351
265
|
</p>
|
|
352
266
|
|
|
353
|
-
|
|
354
|
-
|
|
355
|
-
### 3. Multi-Agent Parallel Swarm Topology
|
|
356
|
-
Run concurrent agents across parallel worktree slots with AST/JSON 3-way merging:
|
|
357
|
-
|
|
358
|
-
<br/>
|
|
359
|
-
|
|
267
|
+
### 3. Polyglot Monorepo Scoped Boundary Resolver
|
|
360
268
|
<p align="center">
|
|
361
|
-
<img src="docs/assets/
|
|
269
|
+
<img src="docs/assets/monorepo-resolver.svg" alt="Polyglot Monorepo Scoped Boundary Resolver" width="100%" />
|
|
362
270
|
</p>
|
|
363
271
|
|
|
364
|
-
|
|
365
|
-
|
|
366
|
-
### 4. Model Context Protocol (MCP) Integration
|
|
367
|
-
Native stdio server exposing task dispatch, gate verification, and risk auditing to client tools (Antigravity, Claude, Cursor):
|
|
368
|
-
|
|
369
|
-
<br/>
|
|
370
|
-
|
|
272
|
+
### 4. Multi-Agent Parallel Swarm Topology
|
|
371
273
|
<p align="center">
|
|
372
|
-
<img src="docs/assets/
|
|
274
|
+
<img src="docs/assets/swarm-topology.svg" alt="Multi-Agent Parallel Swarm Topology" width="100%" />
|
|
373
275
|
</p>
|
|
374
276
|
|
|
375
|
-
|
|
376
|
-
|
|
377
|
-
---
|
|
378
|
-
|
|
379
|
-
<br/>
|
|
380
|
-
|
|
381
|
-
<a id="cli-docs"></a>
|
|
382
|
-
## π οΈ CLI Command Reference (`agentctl`)
|
|
383
|
-
|
|
384
|
-
`agentctl` is the unified command-line interface for `jules-orchestrator-kit`, available via `bin/agentctl.mjs` or `npx jules-orchestrator-kit <command>`.
|
|
385
|
-
|
|
386
|
-
<br/>
|
|
387
|
-
|
|
388
|
-
| Command | Usage | Description | Exit Codes |
|
|
389
|
-
| :--- | :--- | :--- | :--- |
|
|
390
|
-
| `init` | `agentctl init [--interactive] [--tier pro]` | Interactive onboarding wizard & stack oracle inspector generating `.agent/config.yml`. | `0` (Created) |
|
|
391
|
-
| `task create` | `agentctl task create [--title <t>] [--prompt <p>] [--template <id>] [--role <name>] [--tier fast\|complex] [--depends-on <id,...>]` | Interactively authors & scopes falsifiable task envelopes with secret scrubbing, preflight gate checks, specialist role resolution, DAG dependency wiring, and an optional Cost Router tier override. | `0` (Queued), `1` (Unfalsifiable / Secret leak) |
|
|
392
|
-
| `task template` | `agentctl task template [<id>] [--list] [--json]` | Lists and synthesizes specialized web task envelopes (`web-cwv`, `web-wcag`, `web-seo`, `web-playwright`, `web-flaky-heal`, `web-i18n`, `web-ai-access`). | `0` (Synthesized/Listed) |
|
|
393
|
-
| `task optimize` | `agentctl task optimize "<prompt>" [--fix] [--web] [--json]` | Linter & optimizer injecting Google Labs 3-phase exploration budgets, critic steering, and web oracles. | `0` (Scored/Fixed) |
|
|
394
|
-
| `test-gen` | `agentctl test-gen --title <t> --spec <s> [--run]` | Scaffolds falsifiable unit tests, verifies **RED** failure state, and locks test in `scope.deny`. | `0` (Scaffolded/Red) |
|
|
395
|
-
| `rollback` | `agentctl rollback [sessionId \| --latest]` | Restores exact commit, uncommitted files, and cleans orphan task worktrees from pre-flight checkpoints. | `0` (Restored), `1` (Error) |
|
|
396
|
-
| `resume` | `agentctl resume <sessionId> --response "<reply>"` | Streams engineer response back into active Google Jules warm session context window. | `0` (Resumed), `1` (Error) |
|
|
397
|
-
| `dispatch` | `agentctl dispatch --title <t> --prompt <p> [--role <name>] [--tier fast\|complex]` | Dispatches a single task to an AI agent in an isolated worktree, optionally binding a specialist role prompt (`overseer`\|`bolt`\|`sentinel`\|`janitor`) and/or overriding the Cost Router tier. | `0` (Success), `1` (Arg error), `2` (429 Rate limit), `3` (Scope deny), `4` (OODA exhausted), `5` (Diff > 75KB), `6` (Secret leak) |
|
|
398
|
-
| `doctor` | `agentctl doctor [--interactive] [--fix safe]` | Diagnostic DAG check runner & automated transactional repair planner. | `0` (Healthy) |
|
|
399
|
-
| `queue` | `agentctl queue [--interactive] [--dag] [--concurrency <n>] [--json]` | Consumes, inspects, and executes task envelopes in `.agent/jules-queue/`; `--dag` resolves inter-task dependencies via Kahn's algorithm with cycle detection instead of linear FIFO order (supports `--json`). | `0` (Complete) |
|
|
400
|
-
| `swarm` | `agentctl swarm [--interactive] [--json]` | Runs parallel multi-agent swarm across worker slots with process PID liveness detection (supports `--json`). | `0` (Complete) |
|
|
401
|
-
| `scan` | `agentctl scan` | Scans codebase for TODO/FIXME annotations to seed task authoring. | `0` (Scanned) |
|
|
402
|
-
| `review-repair`| `agentctl review-repair <pr-comments.json>`| Parses GitHub PR review comments and synthesizes actionable OODA repair tasks. | `0` (Parsed), `1` (Missing file) |
|
|
403
|
-
| `dashboard` | `agentctl dashboard [port]` | Starts zero-dependency local HTTP telemetry and audit visualizer dashboard. | `0` (Running) |
|
|
404
|
-
| `gate` / `audit`| `agentctl gate --mode working-tree [--json]` | Runs security, secret scanning, and verification gate against working tree or branch (supports `--json`). | `0` (Approved), `3` (Scope violation), `5` (Diff limit), `6` (Secret leak) |
|
|
405
|
-
| `bootstrap` | `agentctl bootstrap [--force] [--json]` | Inspects an untested repository and synthesizes `.agent/config.yml` with a zero-test verification oracle (`php -l`, `compileall`, `dotnet build`, `tsc`, `smoke`). | `0` (Bootstrapped / Existing) |
|
|
406
|
-
| `lock` | `agentctl lock <acquire\|release\|status>`| Manages VFS mutex locks for multi-agent non-overlapping file ownership. | `0` (Locked/Released), `1` (Conflict) |
|
|
407
|
-
| `clean` | `agentctl clean` | Prunes stale git worktrees, lockfiles, and temporary ledgers. | `0` (Clean) |
|
|
408
|
-
| `evidence` | `agentctl evidence <generate\|verify\|show> [--manifest <path>] [--json]` | Generates, verifies, or prints a SHA-256 cryptographic evidence manifest (changed-file hashes + test-file tamper lock) for audit trails. | `0` (Verified/Generated), `1` (Tamper detected / Verification failed) |
|
|
409
|
-
| `escalate` | `agentctl escalate [<sessionId>] [--status] [--flush] [--clear]` | Dispatches or manages webhook escalation incidents across Slack and Discord with Type III Silence Governor and interruption budgeting. | `0` (Dispatched/Buffered) |
|
|
410
|
-
| `flaky` | `agentctl flaky <status\|heal\|reset> [--dispatch] [--dry-run] [--role <name>]` | Manages Wilson-quarantined tests (Exit Code 8) and dispatches automated anti-flakiness healing swarm without assertion weakening. | `0` (Healed/Queued/Listed) |
|
|
411
|
-
| `mcp` | `agentctl mcp` | Starts stdio Model Context Protocol (MCP) server for tool integration. | `0` / Stdio stream |
|
|
412
|
-
| `mcp init` | `agentctl mcp init [--target cursor\|vscode\|claude\|all]` | 1-click scaffolding for Cursor (`.cursor/mcp.json`), VS Code tasks (`tasks.json`), and Claude Desktop. | `0` (Scaffolded) |
|
|
413
|
-
| `version` | `agentctl version` | Outputs orchestrator kit semantic version. | `0` |
|
|
277
|
+
</details>
|
|
414
278
|
|
|
415
279
|
<br/>
|
|
416
280
|
|
|
417
|
-
|
|
281
|
+
<details>
|
|
282
|
+
<summary><b>Multi-Provider Failover & Cost Router SDK</b></summary>
|
|
418
283
|
|
|
419
284
|
<br/>
|
|
420
285
|
|
|
421
|
-
|
|
422
|
-
## π Provider Integration & SDK Usage
|
|
423
|
-
|
|
424
|
-
`jules-orchestrator-kit` supports standard AI agent platforms and programmatic failover routing:
|
|
425
|
-
|
|
426
|
-
### 1. Google Jules Native Integration
|
|
427
|
-
Dispatch tasks using canonical task envelopes or the Google Jules REST v1alpha API.
|
|
428
|
-
|
|
429
|
-
### 2. Multi-Provider Failover SDK (`createFailoverProvider`)
|
|
430
|
-
Programmatically configure ordered provider failover (e.g. falling back to secondary providers on HTTP 429 rate limits):
|
|
431
|
-
|
|
286
|
+
### Multi-Provider Failover SDK (`createFailoverProvider`)
|
|
432
287
|
```javascript
|
|
433
288
|
import { createFailoverProvider, loadConfig } from "jules-orchestrator-kit";
|
|
434
289
|
|
|
@@ -441,9 +296,7 @@ const result = await provider.dispatch(
|
|
|
441
296
|
);
|
|
442
297
|
```
|
|
443
298
|
|
|
444
|
-
###
|
|
445
|
-
Opt-in, config-driven routing between a cheap/fast provider and your primary provider β see [Configuration Reference](#configuration) for the `router:` block. Programmatic usage mirrors `createFailoverProvider`:
|
|
446
|
-
|
|
299
|
+
### Cost Router SDK (`resolveRoutedProvider`)
|
|
447
300
|
```javascript
|
|
448
301
|
import { resolveRoutedProvider, loadConfig } from "jules-orchestrator-kit";
|
|
449
302
|
|
|
@@ -455,41 +308,33 @@ const { provider, classification } = resolveRoutedProvider(
|
|
|
455
308
|
console.log(classification.tier); // "fast" | "complex"
|
|
456
309
|
```
|
|
457
310
|
|
|
458
|
-
|
|
459
|
-
|
|
460
|
-
### 4. Model Context Protocol (MCP) Server
|
|
461
|
-
Expose orchestrator gates and queue controls over stdio to client tools (Antigravity, Claude, Cursor):
|
|
462
|
-
```bash
|
|
463
|
-
npx jules-orchestrator-kit mcp
|
|
464
|
-
```
|
|
311
|
+
</details>
|
|
465
312
|
|
|
466
313
|
<br/>
|
|
467
314
|
|
|
468
|
-
|
|
315
|
+
<details>
|
|
316
|
+
<summary><b>Feature Roadmap & Shipped Milestones</b></summary>
|
|
469
317
|
|
|
470
318
|
<br/>
|
|
471
319
|
|
|
472
|
-
|
|
473
|
-
## πΊοΈ Feature Roadmap & Release History
|
|
474
|
-
|
|
475
|
-
| Feature | Module / Command | Architectural Description | Target Release |
|
|
320
|
+
| Feature | Module / Command | Architectural Description | Status |
|
|
476
321
|
| :--- | :--- | :--- | :---: |
|
|
477
|
-
| **
|
|
478
|
-
| **
|
|
322
|
+
| **Queue Runner Fidelity** | `src/dag-engine.mjs`, `src/engine.mjs` | Queue selection is by task shape rather than file extension, so manifests and READMEs are skipped instead of dispatched, and `--dry-run` leaves the queue untouched. | **v0.38.2** *(Shipped)* |
|
|
323
|
+
| **Release Gate Enforcement & Wizard Smoke Test** | `.github/workflows/jules-audit.yml`, `scripts/release.mjs`, `test/wizard-smoke.test.mjs` | Doc-sync gate runs in CI rather than by hand, releases block on a green CI matrix for `HEAD`, per-test deadlines turn a hang into a failure, and the real `init` wizard is driven end to end over a fake TTY. | **v0.38.1** *(Shipped)* |
|
|
324
|
+
| **Multi-OS CI Matrix & TUI Hardening** | `scripts/run-tests.mjs`, `src/state.mjs`, `src/git.mjs` | Automated 9-job CI matrix across Linux, macOS, and Windows on Node 20/22/24 with raw-mode TUI resilience and native Windows command quoting. | **v0.38.0** *(Shipped)* |
|
|
325
|
+
| **Base64 Secret Detection & Budget Fix** | `src/security.mjs`, `src/budget.mjs` | Secret scanner decodes base64 before matching structured patterns (K8s secrets), and `budget reset` preserves confirmed provider sessions. | **v0.37.0** *(Shipped)* |
|
|
326
|
+
| **Universal AI Crawler Policy & llms.txt** | `src/web-templates.mjs` (`web-ai-access`) | Cross-surface consistency for crawler directives (`robots.txt`, meta tags, `X-Robots-Tag`) and `llms.txt` local route integrity. | **v0.36.0** *(Shipped)* |
|
|
327
|
+
| **Silence Governor & Flaky Test Swarm** | `src/webhook.mjs`, `src/flaky-ledger.mjs` | Notification alert throttling with interruption budgeting, and automated anti-flakiness swarm coordinator. | **v0.35.0** *(Shipped)* |
|
|
328
|
+
| **Rolling 24h Quota & Plan Concurrency** | `src/state.mjs`, `src/config.mjs` | Rolling 24-hour quota accounting matching vendor reset windows and true concurrency limits (3/15/60). | **v0.34.0** *(Shipped)* |
|
|
329
|
+
| **Cost Router & Guided First Run** | `src/router.mjs`, `src/ops/next-step.mjs` | Heuristic task classifier routing trivial tasks to fast models, and guided single-command first run workflow. | **v0.33.0** *(Shipped)* |
|
|
330
|
+
| **DAG Task Queue & Specialist Roles** | `src/dag-engine.mjs`, `src/evidence.mjs` | Kahn's-algorithm dependency queue execution (`queue --dag`), specialist role prompts (`overseer`, `bolt`, `sentinel`, `janitor`), and SHA-256 evidence manifests. | **v0.32.5** *(Shipped)* |
|
|
479
331
|
| **Warm Session Resumption & PR Bundler** | `src/provider.mjs`, `src/engine.mjs` | Multi-turn warm session context streaming via `POST /v1alpha/sessions/{id}:sendMessage` & evidence PR descriptions. | **v0.31.0** *(Shipped)* |
|
|
480
332
|
| **TDD Harness & Prompt Falsifiability Linter** | `agentctl test-gen`, `agentctl task optimize` | Automated RED-state test generator, `scope.deny` test locking, and prompt testability linter with fuzzy path resolution. | **v0.31.0** *(Shipped)* |
|
|
481
333
|
| **Atomic Git Checkpoint & Rollback** | `agentctl rollback` (`src/ops/checkpoint.mjs`) | Pre-flight git HEAD/stash snapshotting, atomic rollback restoration, and 10-session pruning rotation. | **v0.31.0** *(Shipped)* |
|
|
482
|
-
| **Verification Sandbox & SSR Hydration Prober** | `verify.server`, `verify.setup`/`teardown` | Isolated process group dev server probing, Next.js/React SSR panic detection, and deterministic DB hooks. | **v0.31.0** *(Shipped)* |
|
|
483
|
-
| **AST Selective Testing & Escalation Bridge** | `src/dag-engine.mjs`, `agentctl escalate` | Downstream import test resolution, Slack/Discord webhook alerts, and async `agentctl resume` unblocking. | **v0.31.0** *(Shipped)* |
|
|
484
|
-
| **IDE Native MCP Config Scaffolder** | `agentctl mcp init` (`src/ops/ide-scaffold.mjs`) | 1-click scaffolding for Cursor (`.cursor/mcp.json`), VS Code tasks (`tasks.json`), and Claude Desktop. | **v0.31.0** *(Shipped)* |
|
|
485
334
|
| **Interactive UX Engine & TUI Engine** | `src/ux/` (`capabilities`, `key-decoder`, `renderer`, `layout`, `widgets`) | Zero-dependency terminal capabilities detector, sequence key decoder, virtual frame renderer, and widgets. | **v0.30.0** *(Shipped)* |
|
|
486
|
-
| **
|
|
487
|
-
|
|
488
|
-
|
|
489
|
-
| **Onboarding & Stack Oracle Wizard** | `agentctl init --interactive` (`src/wizard-init.mjs`) | Zero-dependency interactive CLI wizard auto-detecting verification oracles, quota tiers, and preset workflows. | **v0.29.0** *(Shipped)* |
|
|
490
|
-
| **Guided Task Authoring Subsystem** | `agentctl task create` (`src/wizard-task.mjs`) | Guided task authoring with TODO candidate harvesting, Shannon entropy secret scrubbing, and guardrail footer synthesis. | **v0.29.0** *(Shipped)* |
|
|
491
|
-
| **P0 Remediation & Safety Alignment** | Queue, Task Envelope & Secrets (`src/wizard-task.mjs`) | Canonical queue path alignment (`.agent/jules-queue/`), path traversal guards, atomic writes, multiline secret scans, and JSON headers. | **v0.29.1** *(Shipped)* |
|
|
492
|
-
| **PR Review Auto-Remediation Loop** | `agentctl review-repair` (`src/review-repair.mjs`) | Ingests GitHub PR review comments (`CHANGES_REQUESTED`), extracts line/file context, and dispatches automated OODA repair turns. | **v0.27.0** *(Shipped)* |
|
|
335
|
+
| **PR Review Auto-Remediation Loop** | `agentctl review-repair` (`src/review-repair.mjs`) | Ingests GitHub PR review comments (`CHANGES_REQUESTED`), extracts line/file context, and dispatches automated repair turns. | **v0.27.0** *(Shipped)* |
|
|
336
|
+
|
|
337
|
+
</details>
|
|
493
338
|
|
|
494
339
|
<br/>
|
|
495
340
|
|
|
@@ -497,9 +342,9 @@ npx jules-orchestrator-kit mcp
|
|
|
497
342
|
|
|
498
343
|
<br/>
|
|
499
344
|
|
|
500
|
-
## π Documentation &
|
|
345
|
+
## π Documentation & External References
|
|
501
346
|
|
|
502
|
-
- [**System Architecture & Pipeline Overview**](./docs/architecture.md) β Comprehensive technical sequence
|
|
347
|
+
- [**System Architecture & Pipeline Overview**](./docs/architecture.md) β Comprehensive technical sequence diagrams and control plane specifications.
|
|
503
348
|
- [**Google Jules Official Documentation**](https://jules.google) β Official platform overview and API specifications for Google Jules.
|
|
504
349
|
- [**Examples & Task Envelope Recipes**](./EXAMPLES.md) β Production YAML and Markdown task envelopes.
|
|
505
350
|
- [**Changelog**](./CHANGELOG.md) β Full release history and migration guides.
|
|
@@ -510,6 +355,17 @@ npx jules-orchestrator-kit mcp
|
|
|
510
355
|
|
|
511
356
|
<br/>
|
|
512
357
|
|
|
358
|
+
## βοΈ Disclaimer
|
|
359
|
+
|
|
360
|
+
`jules-orchestrator-kit` is an independent, community-driven open-source project and is not affiliated with, endorsed by, or sponsored by Google, Google LLC, or Alphabet Inc. "Google", "Google Jules", and related marks are trademarks of Google LLC.
|
|
361
|
+
|
|
362
|
+
<br/>
|
|
363
|
+
|
|
364
|
+
---
|
|
365
|
+
|
|
366
|
+
<br/>
|
|
367
|
+
|
|
513
368
|
<div align="center">
|
|
514
|
-
<p><b>jules-orchestrator-kit</b> β’ Built with zero external dependencies for Google Jules and
|
|
369
|
+
<p><b>jules-orchestrator-kit</b> β’ Built with zero external dependencies for Google Jules and autonomous agent workflows.</p>
|
|
515
370
|
</div>
|
|
371
|
+
|