agent-nuvira 1.16.1 → 1.17.0

This diff represents the content of publicly available package versions that have been released to one of the supported registries. The information contained in this diff is provided for informational purposes only and reflects changes between package versions as they appear in their respective public registries.
Files changed (2) hide show
  1. package/README.md +187 -23
  2. package/package.json +1 -1
package/README.md CHANGED
@@ -47,8 +47,16 @@ agent-nuvira config list
47
47
  - **Security scan CLI** — `buff security scan` detects PII, prompt injections, and dangerous code patterns
48
48
  - **Feedback & rating system** — `buff feedback record/list/stats/clear` drives self-improvement scoring
49
49
  - **Marketplace unified CLI** — `buff marketplace browse/search/install/info` for workflow templates + plugins
50
- - **MCP (Model Context Protocol) integration** — connect to databases, APIs, and file systems via MCP servers
50
+ - **MCP (Model Context Protocol) integration** — connect to databases, APIs, and file systems via MCP servers with SSE transport support
51
51
  - **AST-aware code editing** — structural analysis engine understands functions, classes, methods across JS/TS/Python/Go/Rust
52
+ - **Auto error-repair engine** — automatic diagnosis and repair of test failures with configurable retry budgets
53
+ - **A2A (Agent-to-Agent) Protocol** — inter-agent communication standard for multi-machine collaboration
54
+ - **CI/CD headless mode** — `buff ci` for automated pipelines with GitHub Actions integration
55
+ - **npm publishing & one-line install** — `npx agent-nuvira` and `npx buff` for zero-setup onboarding
56
+ - **Interactive development mode** — `buff execute` without a goal launches a guided interactive loop with session save/resume, follow-up suggestions, and failure analysis
57
+ - **Session persistence** — save and resume development sessions across CLI restarts with full history
58
+ - **Failure analysis** — automatic diagnosis of agent failures with specific recovery options per agent type
59
+ - **Follow-up suggestions** — LLM-powered contextual next-step recommendations after goal completion
52
60
  - **Configuration** via JSON config file + environment variables
53
61
  - **No server dependency** — no telemetry, no subscriptions, no outbound calls to a hosted backend
54
62
 
@@ -897,17 +905,35 @@ agent-nuvira execute "migrate database schema" --context-prune aggressive
897
905
  The pruner automatically compresses the shared agent context between pipeline steps using 5 strategies: metadata stripping, file change collapsing, conversation truncation, artifact summarization, and aggressive fallback.
898
906
 
899
907
 
900
- The pipeline runs these agents in sequence (with parallelization where possible):
901
- 1. **Planner** — Analyzes the goal, creates a task plan
902
- 2. **Context Gatherer** — Scans the codebase for relevant files
903
- 3. **Writer** — Implements the code changes
904
- 4. **Reviewer** — Validates the changes for bugs and style (optional)
905
- 5. **Tester** — Runs tests in a sandbox (optional)
906
- 6. **Runner** — Executes the program to verify it works (optional)
907
- 7. **Debugger** — Iterates on test failures (optional)
908
- 8. **Git Agent** — Commits changes to a branch (optional)
909
- 9. **Package Agent** — Bumps versions and generates changelogs (optional)
910
- 10. **GitHub Release Agent** — Creates tags and releases (optional)
908
+ The pipeline runs these agents in dependency-aware order with parallelization:
909
+
910
+ | # | Agent | Type | Description |
911
+ |---|-------|------|-------------|
912
+ | 1 | **Planner** | Core | Analyzes the goal, creates a dependency-aware task plan |
913
+ | 2 | **Context Gatherer** | Core | Scans the codebase for relevant files and artifacts |
914
+ | 3 | **Security Scanner** | Safety | Scans for PII, prompt injection, and dangerous patterns |
915
+ | 4 | **Writer** | Core | Implements the code changes based on plan and context |
916
+ | 5 | **Reviewer** | Quality | Validates changes for bugs, security, and code style |
917
+ | 6 | **Tester** | Testing | Runs tests in a sandboxed temp directory or Docker container |
918
+ | 7 | **Debugger** | Testing | Iteratively diagnoses and fixes test failures via LLM |
919
+ | 8 | **Runner** | Execution | Executes shell commands to verify the program works |
920
+ | 9 | **MCP Agent** | Integration | Invokes external tools from connected MCP servers |
921
+ | 10 | **Skill Runner** | Learning | Executes compiled skill scripts as pre-built task plans |
922
+ | 11 | **Git Agent** | Publishing | Creates branches, commits with LLM-generated messages |
923
+ | 12 | **PR Description** | Publishing | Generates PR descriptions from git diff via LLM |
924
+ | 13 | **Package Agent** | Publishing | Bumps versions, builds, publishes to npm |
925
+ | 14 | **GitHub Release** | Publishing | Creates tags, release notes, and GitHub releases |
926
+
927
+ **Parallel execution:** Independent agents (e.g., Reviewer + Tester) run concurrently via `Promise.all()`. Exclusive agents (Runner, Debugger) get dedicated access. Results are merged with conflict resolution.
928
+
929
+ **Interactive development mode:** `buff execute` without a goal launches an interactive loop with:
930
+ - **Model picker** — Choose your provider/model interactively
931
+ - **Session tracking** — Full history of goals executed in the session
932
+ - **Failure analysis** — Per-agent-type diagnosis with recovery actions
933
+ - **Follow-up suggestions** — LLM-powered contextual next steps
934
+ - **/fix** — Retry the last failed goal with failure context
935
+ - **/save / /resume** — Save and restore sessions across restarts
936
+ - **/suggest** — Search past trajectories for similar goals
911
937
 
912
938
  ---
913
939
 
@@ -933,15 +959,20 @@ Adapter Adapter Adapter Adapter Adapter
933
959
  │ Core Pipeline │
934
960
  │ ┌────────────────────────┐ │
935
961
  │ │ Orchestrator │ │
936
- │ │ ├─ Planner │ │
937
- │ │ ├─ ContextGather │ │
938
- │ │ ├─ Writer │ │
939
- │ │ ├─ Reviewer │ │
940
- │ │ ├─ Tester │ │
941
- │ │ ├─ Runner │ │
942
- │ │ ├─ Debugger │ │
943
- │ │ ├─ GitAgent │ │
944
- │ │ └─ SkillRunner │ │
962
+ │ │ │ ├─ Planner │ │
963
+ │ │ ├─ ContextGatherer │ │
964
+ │ │ ├─ Writer │ │
965
+ │ │ ├─ Reviewer │ │
966
+ │ │ ├─ Tester │ │
967
+ │ │ ├─ Runner │ │
968
+ │ │ ├─ Debugger │ │
969
+ │ │ ├─ GitAgent │ │
970
+ │ │ ├─ PackageAgent │ │
971
+ │ │ ├─ GitHubReleaseAgent │ │
972
+ │ │ ├─ SecurityAgent │ │
973
+ │ │ ├─ SkillRunner │ │
974
+ │ │ ├─ MCPAgent │ │
975
+ │ │ └─ PRDescriptionAgent │ │
945
976
  │ └────────────────────────┘ │
946
977
  │ │
947
978
  │ ┌────────────────────────┐ │
@@ -1322,12 +1353,145 @@ npx tsc --noEmit
1322
1353
  | 3.9 | Security scan CLI (`buff security scan`) | ✅ Complete |
1323
1354
  | 3.10 | Feedback & rating system (`buff feedback`) | ✅ Complete |
1324
1355
  | 3.11 | Marketplace unified CLI (`buff marketplace browse/search/install`) | ✅ Complete |
1325
- | **Phase 4: Industry Standards** | *(in progress)* | |
1326
- | 4.1 | MCP (Model Context Protocol) integration — MCP client/manager/CLI | ✅ Complete |
1356
+ | **Phase 4: Industry Standards** | | |
1357
+ | 4.1 | MCP (Model Context Protocol) — client/manager/CLI with SSE transport + Firecrawl | ✅ Complete |
1327
1358
  | 4.2 | AST-aware code editing — structural analysis engine (JS/TS/Python/Go/Rust) | ✅ Complete |
1359
+ | 4.3 | Auto error-repair engine — diagnosis & retry budgets for test failures | ✅ Complete |
1360
+ | 4.4 | A2A (Agent-to-Agent) Protocol — inter-agent communication standard | ✅ Complete |
1361
+ | 4.5 | CI/CD headless mode — `buff ci` with GitHub Actions integration | ✅ Complete |
1362
+ | 4.6 | npm publishing & one-line install — `npx agent-nuvira` / `npx buff` | ✅ Complete |
1363
+ | **Phase 5: Interactive UX** | | |
1364
+ | 5.1 | Interactive dev mode — guided loop with model picker, session save/resume | ✅ Complete |
1365
+ | 5.2 | Failure analysis — per-agent-type diagnosis with recovery actions | ✅ Complete |
1366
+ | 5.3 | Follow-up suggestions — LLM-powered contextual next-step recommendations | ✅ Complete |
1367
+ | 5.4 | /fix command — retry last failed goal with failure context | ✅ Complete |
1368
+ | 5.5 | Test coverage — 1830+ tests across 55 test files | ✅ Complete |
1328
1369
 
1329
1370
  ---
1330
1371
 
1372
+ ## Version History
1373
+
1374
+ | Version | Date | Key Changes |
1375
+ |---------|------|-------------|
1376
+ | **v1.0.0** | Apr 2026 | Initial release — Core CLI with chat, 5 providers, config, models |
1377
+ | **v1.1.0** | Apr 2026 | Model discovery with search/filter |
1378
+ | **v1.2.0** | Apr 2026 | AI-assisted file editing (edit command) |
1379
+ | **v1.3.0** | May 2026 | Implementation plans (plan command) |
1380
+ | **v1.4.0** | May 2026 | Multi-agent pipeline (execute command) with Planner, Writer, ContextGatherer |
1381
+ | **v1.5.0** | May 2026 | Additional agents — Tester, Runner, Debugger |
1382
+ | **v1.6.0** | Jun 2026 | Agent retry logic, format validation, git integration |
1383
+ | **v1.7.0** | Jun 2026 | Phase 1 features — plugin system, cost tracking, logging |
1384
+ | **v1.8.0** | Jun 2026 | Native embeddings, vector store, trajectory memory |
1385
+ | **v1.9.0** | Jul 2026 | Workflow templates, model benchmarking |
1386
+ | **v1.10.0** | Jul 2026 | Docker sandbox, provider health dashboard |
1387
+ | **v1.11.0** | Jul 2026 | Skill compiler, context pruner, model switching |
1388
+ | **v1.12.0** | Jul 2026 | VS Code extension, web dashboard, agent federation |
1389
+ | **v1.13.0** | Jul 2026 | Hybrid model routing, team collaboration, Agent SDK |
1390
+ | **v1.14.0** | Jul 2026 | Provider fallback, security scan, feedback system, marketplace CLI |
1391
+ | **v1.14.6** | Jul 2026 | Skill compiler system, context-window pruner, Docker Compose onboarding |
1392
+ | **v1.15.0** | Aug 2026 | npm publishing — `npx buff` / `npx agent-nuvira` live on npm (1.3 MB) |
1393
+ | **v1.15.1** | Aug 2026 | Interactive dev mode — model picker, session tracking, /save / /resume, /suggest |
1394
+ | **v1.15.2** | Aug 2026 | Windows compatibility fixes |
1395
+ | **v1.15.3** | Aug 2026 | Accessibility fix — `window.open` → native `<a>` tags |
1396
+ | **v1.15.4** | Aug 2026 | Search/filter bar, column count toggle, speech provider section |
1397
+ | **v1.15.5** | Aug 2026 | SSE header support for MCP |
1398
+ | **v1.15.6** | Aug 2026 | Firecrawl integration for web search |
1399
+ | **v1.16.0** | Aug 2026 | Comprehensive MCP README docs, SSE header support |
1400
+ | **v1.16.1** | Aug 2026 | Interactive dev mode enhancements — failure analysis, follow-up suggestions, /fix command, 35 new unit tests |
1401
+
1402
+ ---
1403
+
1404
+ ## Phase-Wise Feature Summary
1405
+
1406
+ ### Phase 0: Foundation — Core CLI & Provider Layer
1407
+ | Feature | Description |
1408
+ |---------|-------------|
1409
+ | **5 Inference Providers** | Groq, NVIDIA NIM, Google Gemini, OpenRouter, Local (Ollama/HuggingFace/GGML) |
1410
+ | **Unified CLI** | 25+ commands via Commander.js with shared options |
1411
+ | **Config System** | JSON config file + env vars + CLI flags priority chain |
1412
+ | **Streaming** | Real-time token-by-token output for all 5 providers |
1413
+ | **Response Caching** | SQLite-backed cache with configurable TTL |
1414
+ | **Chat Interface** | Interactive chat with conversation history and `/` commands |
1415
+ | **File Editing** | AI-assisted file editing with dry-run mode |
1416
+ | **Implementation Plans** | Codebase-aware plan generation with architecture impact analysis |
1417
+
1418
+ ### Phase 1: Quick Wins — Developer Experience
1419
+ | Feature | Description |
1420
+ |---------|-------------|
1421
+ | **Plugin System** | Programmatic API + auto-discovery from `~/.buff/plugins/` |
1422
+ | **Project Scaffolding** | `buff init` with 5 built-in templates + interactive provider wizard |
1423
+ | **Model Discovery** | `buff models` with search/filter across all providers |
1424
+ | **Model Switching** | Context-preserving provider/model switch mid-session |
1425
+ | **Cost Tracking** | Per-provider, per-session, and monthly cost dashboards |
1426
+ | **History Search** | Keyword + semantic search across past conversations |
1427
+ | **Skill Compiler** | Auto-extracts reusable patterns from trajectories into runnable skills |
1428
+ | **Context Pruner** | 5-strategy token compression for long agent chains |
1429
+
1430
+ ### Phase 2: Structural Changes — Memory & Infrastructure
1431
+ | Feature | Description |
1432
+ |---------|-------------|
1433
+ | **Vector Store** | Cosine similarity search over embedded trajectories |
1434
+ | **Trajectory Store** | Few-shot example storage with quality scoring |
1435
+ | **3-Tier Embedder** | Xenova (fast) → Python (medium) → LLM (fallback) |
1436
+ | **Workflow Marketplace** | 10 built-in templates + GitHub registry with install/publish |
1437
+ | **Model Benchmarking** | 21 standardized coding tasks with scoring and A/B comparison |
1438
+ | **Docker Sandbox** | 8 base images, resource limits, network-isolated execution |
1439
+ | **Provider Health** | `buff doctor` with color-coded status, watch mode, auto-fix |
1440
+ | **Memory Compression** | Automatic trajectory summarization with configurable retention |
1441
+
1442
+ ### Phase 3: Major Upgrades — Advanced Agent Systems
1443
+ | Feature | Description |
1444
+ |---------|-------------|
1445
+ | **VS Code Extension** | 9 commands, inline code suggestions, diff viewer, agent progress panel |
1446
+ | **Agent Federation** | Multi-machine collaboration via A2A protocol, server, and client |
1447
+ | **Web Dashboard** | React + Recharts + DAG visualization, model health, cost charts |
1448
+ | **Hybrid Model Routing** | Complexity-based model selection with cost optimization |
1449
+ | **Team Collaboration** | Git-synced shared config, memory, and review pipelines |
1450
+ | **Agent SDK** | `@agent-nuvira/sdk` npm package with scaffolding CLI |
1451
+ | **Provider CLI** | `buff provider list/health` with per-provider diagnostics |
1452
+ | **Provider Fallback** | Auto-failover with circuit breaker and configurable chain |
1453
+ | **Security Scanner** | Detects PII, prompt injections, and dangerous code patterns |
1454
+ | **Feedback System** | `buff feedback record/list/stats/clear` drives self-improvement |
1455
+ | **Marketplace CLI** | Unified `buff marketplace browse/search/install/info` |
1456
+
1457
+ ### Phase 4: Industry Standards — Protocol & Integration
1458
+ | Feature | Description |
1459
+ |---------|-------------|
1460
+ | **MCP Protocol** | Model Context Protocol client/manager with stdio + SSE transport |
1461
+ | **AST Editing Engine** | Structural code analysis for JS/TS/Python/Go/Rust |
1462
+ | **Auto Error-Repair** | Automatic diagnosis and repair with configurable retry budgets |
1463
+ | **A2A Protocol** | Agent-to-Agent communication standard for federation |
1464
+ | **CI/CD Headless** | `buff ci` for automated pipelines with GitHub Actions |
1465
+ | **npm Publishing** | `npx agent-nuvira` / `npx buff` for zero-setup onboarding |
1466
+
1467
+ ### Phase 5: Interactive UX — Developer Experience
1468
+ | Feature | Description |
1469
+ |---------|-------------|
1470
+ | **Interactive Dev Mode** | Guided loop with model picker, session management, and goal tracking |
1471
+ | **Session Save/Resume** | Save and restore development sessions with full history |
1472
+ | **Failure Analysis** | Per-agent-type diagnosis with specific recovery actions |
1473
+ | **Follow-up Suggestions** | LLM-powered contextual next-step recommendations |
1474
+ | **/fix Command** | Retry last failed goal with failure context |
1475
+ | **Graceful Error Recovery** | Rate-limit handling, auth failures, and network error recovery |
1476
+
1477
+ ### Agent Catalog — 15 Agent Roles & Management
1478
+ | Agent/Component | Type | Description |
1479
+ |-----------------|------|-------------|
1480
+ | **PlannerAgent** | Core | Analyzes goals, creates dependency-aware task plans |
1481
+ | **ContextGathererAgent** | Core | Scans codebase, identifies relevant files and artifacts |
1482
+ | **WriterAgent** | Core | Implements code changes based on plan and gathered context |
1483
+ | **ReviewerAgent** | Core | Validates changes for bugs, security, and style |
1484
+ | **RunnerAgent** | Execution | Executes shell commands and captures output |
1485
+ | **TesterAgent** | Testing | Runs tests in sandboxed temp directory or Docker container |
1486
+ | **DebuggerAgent** | Testing | Iteratively diagnoses and fixes test failures via LLM |
1487
+ | **GitAgent** | Publishing | Creates branches, commits with LLM messages, generates PR descriptions |
1488
+ | **PackageAgent** | Publishing | Bumps version, builds, publishes to npm, generates changelogs |
1489
+ | **GitHubReleaseAgent** | Publishing | Creates tags, release notes, and GitHub releases via `gh` CLI or API |
1490
+ | **SecurityAgent** | Safety | Scans for PII, prompt injection, and dangerous code patterns |
1491
+ | **SkillRunnerAgent** | Learning | Executes compiled skill scripts as pre-built task plans |
1492
+ | **MCPAgent** | Integration | Invokes MCP tools from connected servers via stdio or SSE transport |
1493
+ | **Orchestrator** | Management | Coordinates all agents with dependency-aware scheduling, parallel execution, context pruning, and interactive recovery |
1494
+
1331
1495
  ## License
1332
1496
 
1333
1497
  MIT
package/package.json CHANGED
@@ -1,6 +1,6 @@
1
1
  {
2
2
  "name": "agent-nuvira",
3
- "version": "1.16.1",
3
+ "version": "1.17.0",
4
4
  "description": "Agent-Nuvira: Multi-agent AI coding CLI — plan, write, review, test, and publish code with local models (Ollama) or cloud APIs (Groq, NVIDIA NIM, Google Gemini, OpenRouter)",
5
5
  "type": "module",
6
6
  "main": "dist/index.js",