apt-mcp-agent-setup 3.2.7 → 3.3.1

This diff represents the content of publicly available package versions that have been released to one of the supported registries. The information contained in this diff is provided for informational purposes only and reflects changes between package versions as they appear in their respective public registries.
Files changed (78) hide show
  1. package/bin/cli.js +1 -1
  2. package/bundle/NOTICES.md +0 -0
  3. package/bundle/core.enc +0 -0
  4. package/bundle/mcp-rules.enc +0 -0
  5. package/bundle/skills-aso.enc +0 -0
  6. package/bundle/skills-ba.enc +0 -0
  7. package/bundle/skills-base.enc +0 -0
  8. package/bundle/skills-be.enc +0 -0
  9. package/bundle/skills-design.enc +0 -0
  10. package/bundle/skills-fe.enc +0 -0
  11. package/bundle/skills-mobile.enc +0 -0
  12. package/bundle/skills-pm.enc +0 -0
  13. package/integrity-manifest.json +78 -65
  14. package/package.json +1 -1
  15. package/release-check.js +3 -1
  16. package/src/backends/browser-tools.js +1 -1
  17. package/src/backends/index-command.js +1 -1
  18. package/src/backends/index-store.js +1 -1
  19. package/src/backends/memory-2026.8.31.json +34 -0
  20. package/src/backends/profiles.js +1 -1
  21. package/src/core/admission-fence.js +1 -0
  22. package/src/core/aso-import.js +1 -1
  23. package/src/core/catalog.js +1 -1
  24. package/src/core/codex-host-metadata.js +1 -1
  25. package/src/core/codex-transcript-receipt.js +1 -0
  26. package/src/core/context-metrics.js +1 -1
  27. package/src/core/doctor.js +1 -1
  28. package/src/core/execution.js +1 -1
  29. package/src/core/hook-bridge.js +1 -1
  30. package/src/core/host-adapters.js +1 -1
  31. package/src/core/memory-store.js +1 -1
  32. package/src/core/model-inventory.js +1 -1
  33. package/src/core/model-routing.js +1 -1
  34. package/src/core/native-dispatch.js +1 -0
  35. package/src/core/npm-cli.js +1 -1
  36. package/src/core/owned-lock.js +1 -1
  37. package/src/core/presentations.js +1 -1
  38. package/src/core/quality.js +1 -1
  39. package/src/core/run-compatibility.js +1 -0
  40. package/src/core/runtime-info.js +1 -1
  41. package/src/core/session.js +1 -1
  42. package/src/core/skill-names.js +1 -1
  43. package/src/core/workspace-path.js +1 -1
  44. package/src/installer/bounded-command.js +1 -1
  45. package/src/installer/candidate-preflight.js +1 -0
  46. package/src/installer/connection-activation.js +1 -0
  47. package/src/installer/connection-supervisor.js +1 -0
  48. package/src/installer/global-setup.js +1 -1
  49. package/src/installer/host-hooks.js +1 -1
  50. package/src/installer/managed-config.js +1 -1
  51. package/src/installer/platform-config.js +1 -1
  52. package/src/installer/prerequisites.js +1 -1
  53. package/src/installer/presentations.js +1 -1
  54. package/src/installer/project-setup.js +1 -1
  55. package/src/installer/registration-receipt.js +1 -0
  56. package/src/installer/registrations.js +1 -0
  57. package/src/installer/runtime-activation.js +1 -1
  58. package/src/installer/runtime-child.js +1 -0
  59. package/src/installer/runtime-launcher.js +30 -5
  60. package/src/installer/runtime-lock.js +1 -1
  61. package/src/installer/runtime-store.js +1 -1
  62. package/src/installer/runtime-supervisor.js +1 -0
  63. package/src/installer/setup-wizard.js +1 -1
  64. package/src/installer/workspace-update.js +1 -0
  65. package/src/license/crypto.js +1 -1
  66. package/src/license/fingerprint.js +1 -1
  67. package/src/license/terms.js +1 -1
  68. package/src/license/verify.js +1 -1
  69. package/src/presets/index.js +1 -1
  70. package/src/proxy/backends.js +1 -1
  71. package/src/proxy/browser-sessions.js +1 -1
  72. package/src/proxy/core-tools.js +1 -1
  73. package/src/proxy/pipeline.js +1 -1
  74. package/src/proxy/router.config.js +1 -1
  75. package/src/proxy/router.js +1 -1
  76. package/src/proxy/server.js +1 -1
  77. package/src/templates/apt-runtime.md +3 -3
  78. package/src/templates/mcp-tools.md +7 -5
@@ -4,10 +4,10 @@
4
4
  2. Call `route_request` for a new request or changed intent, including target package/stack and explicit skills. Load the selected specialist and companion skills with `skill_load`. Do not read the complete trigger catalog into the model context.
5
5
  3. Read-only questions, reviews and running existing tests do not require a code-change pipeline. Before executable changes, inspect `pipeline_status` and bind `pipeline_start` or `pipeline_use` with canonical task/run identity and an explicit workspace.
6
6
  4. Small changes use scope, fix/check and review. Features use clarification only where needed, a plan with acceptance criteria, implementation, verification, review and acceptance. Bug fixes start with diagnosis and regression coverage. Do not require three questions, a PRD or external issues when the task is already clear.
7
- 5. Use `pipeline_next` to reserve work before dispatch. Delegate useful independent research and reviews to native host subagents, with at most four active actors including the coordinator. Only one writer per workspace. Workers must not start/reset/rebind the parent pipeline.
7
+ 5. Use `pipeline_next` to reserve work before dispatch. Use its pinned `nativeDispatch` fields only where the native tool supports them. Codex model overrides use a fresh brief with `fork_turns:"none"` and the exact `nativeTaskName`; a follow-up cannot change an existing worker's model. Include APT attempt, actor and session markers in the brief. Delegate useful independent research and reviews to native host subagents, with at most four active actors including the coordinator. Only one writer per workspace. Workers must not start/reset/rebind the parent pipeline.
8
8
  6. Pin a quality plan for new executable runs: stable criteria, expected outcomes, concrete check IDs, target platforms and required review roles. Use `pipeline_verify` with a supported test report and named assertions, then `pipeline_checkpoint` with the current evidence IDs. Exit code, build, screenshot or a free-form summary cannot prove functional acceptance. Missing required checks keep the run unaccepted.
9
- 7. Record separate technical and Product/UX review verdicts when the work affects a user flow. Review staged, unstaged and relevant untracked changes. Keep finding IDs open across reviews until a current fix or reasoned dismissal is confirmed by a reviewer. High-risk work needs a host-observed independent technical reviewer; actor names and self-reported models do not prove independence. Rerun checks after affected changes. Review-only requests produce findings, not unsolicited edits or MR comments.
10
- 8. Stop after ten repair/check attempts per work item or three repeated outcomes without progress. Preserve budgets and blocked/paused states across resume. User interruption always wins over continuation hooks.
9
+ 7. Record separate technical and Product/UX review verdicts when the work affects a user flow. Review staged, unstaged and relevant untracked changes. Keep finding IDs open across reviews until a current fix or reasoned dismissal is confirmed by a reviewer. The native reviewer records its own structured verdict with `pipeline_checkpoint(action="review")`; nested work items use `spec`/`quality`, while run-level review uses `technical`/`product`. High-risk work needs host proof of the reviewer and verdict authorship; actor names, worker receipts alone and self-reported models do not prove it. Rerun checks after affected changes. Review-only requests produce findings, not unsolicited edits or MR comments.
10
+ 8. `pipeline_verify(kind="review")` cannot satisfy a review gate and is rejected before command execution. Use the `requiredReview` returned by `pipeline_next`, then record a structured review on the reserved attempt. Stop after ten repair/check attempts per work item or three repeated outcomes without progress. Preserve budgets and blocked/paused states across resume. User interruption always wins over continuation hooks.
11
11
  9. Native tool enforcement and automatic continuation depend on observed host capability. If unavailable, use sequential execution and state that limitation. Never claim independent review without an independent reviewer.
12
12
  10. Keep prose short and clear, reuse existing code, and load only current-step references. Strong caveman style is opt-in. Publishing, merging and external communication require authorization for those actions.
13
13
 
@@ -4,17 +4,19 @@
4
4
 
5
5
  For Figma and Stitch tool names and authentication, read `docs/agents/design-config.md` when working on design tasks.
6
6
 
7
- - Bootstrap: `session_bootstrap(sessionId, contextEpoch, host, capabilities)` identifies this context and reports resumable work. Pass native model capability and account scope only when the host supplies them.
7
+ - Bootstrap: `session_bootstrap(sessionId, contextEpoch, host, capabilities)` identifies this context and reports resumable work. When the native tool schema is visible, describe `capabilities.nativeDispatch` with its tool, selection mode, model/effort field names, fork modes and requestable model IDs. This coordinator declaration permits a request, but is not a host-observed model receipt. An absent capability field means unknown, not a permanent denial.
8
8
  - Route: `route_request(sessionId, requestId, prompt, target, stack, explicitSkills)` returns the workflow, specialist, review roles, platform verification requirements and phase-specific skills. Load current-step skills with `skill_load` and agent/workflow content with `context_load` for each new worker or context epoch.
9
- - Models: `model_list(sessionId, offset, limit, query)` reports inventory source, account scope, coverage and eligibility. CLI metadata is reference only. Use the model decision from `pipeline_next`; compare requested and host-observed model receipts after dispatch.
9
+ - Models: `model_list(sessionId, offset, limit, query)` reports advertised, requestable and eligible models with source, account scope and coverage. CLI and Codex local cache entries are reference only. Use the pinned `nativeDispatch` from `pipeline_next`; compare its selector with the actual spawn arguments, then keep the host-reported worker model separate. An `inherit` selector or a model chosen in the IDE does not confirm a child model.
10
10
  - Start: `pipeline_start(pipeline, task, workspace, acceptanceCriteria, scope, qualityPlan, productImpact, sessionId)` creates a versioned run. Bootstrap the session in the intended workspace first. `workspace` is a schema-valid object, such as `{mode:"current",confirmed:true}`; a string or an unbound global server cannot create a run. On Antigravity, use the workspace MCP alias. New executable runs require at least one acceptance criterion string that exactly matches `qualityPlan.criteria[].expectedOutcome`. The plan uses `criteria: [{id, expectedOutcome, source, checkIds}]`, `checks: [{id, criterionIds, format, testIds, required}]`, and `reviews: [{id, required}]`. A recent user-facing web or mobile route requires its platform check and Product review. Use `pipeline_use(task, runId)` to resume the same run.
11
- - Execute: `pipeline_next(task, runId, sessionId, role, requestId)` reserves an attempt. Include its APT attempt markers in the native worker brief. Native host agents perform the work; the coordinator controls transitions and continues through ready parts of an approved feature without asking for approval after each internal step. Resource and quota waits remain resumable.
12
- - Verify: `pipeline_verify` records authorized argv execution under the run workspace. For functional acceptance, pass `checkId`, supported `reportFormat` and named test IDs in the pinned plan. Zero, skipped or unreadable tests and unrelated successful commands remain unverified. `pipeline_checkpoint` records the latest review verdict per stage and persistent finding IDs through `action=review`; resolution needs current evidence and later reviewer confirmation. A direct start without a bound route needs Technical `impactAssessment` before acceptance. User-flow impact requires Product review and platform checks. Missing required platform checks leave the run unaccepted.
11
+ - Execute: `pipeline_next(task, runId, sessionId, role, requestId)` reserves an attempt. Pass its exact `nativeTaskName`, model/tier, reasoning effort and fork mode to the native tool where that tool supports them. Codex model overrides require `fork_turns:"none"`; put `APT_ATTEMPT_ID`, `APT_ACTOR_ID` and `APT_SESSION_ID` in the fresh worker brief. A `followup_task` keeps the existing worker's model; dispatch a new worker when a new model is required. The coordinator controls transitions. Resource and quota waits remain resumable.
12
+ - Verify: `pipeline_verify` records authorized argv execution under the run workspace. For functional acceptance, pass `checkId`, supported `reportFormat` and named test IDs in the pinned plan. Zero, skipped or unreadable tests and unrelated successful commands remain unverified. `pipeline_verify(kind="review")` is rejected before running a command. Read `pipeline_next.requiredReview`: nested work items require `spec`/`quality`, while run-level gates require `technical`/`product`. The reserved reviewer records its own structured verdict with `pipeline_checkpoint(action="review")`; resolution of a persistent finding needs current evidence and later reviewer confirmation. A direct start without a bound route needs Technical `impactAssessment` before acceptance. User-flow impact requires Product review and platform checks. Missing required platform checks leave the run unaccepted.
13
13
  - Inspect: `pipeline_status` reports current criterion/check status, open findings, reviewer provenance and pending creation intents without binding sessions. `pipeline_reset(task, runId, sessionId, confirm:true, target:"creation")` abandons a matching unfinished intent while preserving its file. The default `target:"run"` cancels a run without deleting history. `pipeline_prepare_merge` only prepares a report.
14
14
  - Read-only questions, review and existing test execution do not require a mutation pipeline. Scope-limited changes use the short pipeline; increased risk adds required checks.
15
15
  - MCP gates cover proxied calls. Native tool gates and continuation require observed host hooks. Plan Mode and user interruption take precedence.
16
16
  - Codegraph failures permit file-search fallback. The proxy refreshes the review graph before impact analysis; it must report failure rather than use stale results.
17
- - `mcp_agent_health` reports process, initialization and functional readiness separately. `npx apt-mcp-agent-setup doctor` performs read-only configuration inspection.
17
+ - `mcp_agent_health` reports process, initialization and functional readiness separately. `npx apt-mcp-agent-setup doctor` performs read-only configuration inspection.
18
+ - Backend catalogs are available without starting child processes. A tool call starts only its target backend. Pass `backend` to `mcp_agent_health` when a live probe is required.
19
+ - Use `npx apt-mcp-agent-setup mcp registrations --json` to inspect workspace aliases and `mcp cleanup --dry-run` before reviewing legacy validation aliases. Cleanup requires a reviewed plan and preserves a backup.
18
20
  - Browser work uses the lazy `browser-local` backend when the preset has it. Pass `aptTask` and `aptWorker` on browser calls. Use snapshot and console evidence. A page load or screenshot alone is not a feature check. After an uncertain browser action, inspect state before retrying and never replay a side-effecting call automatically.
19
21
  - `presentation_prepare` returns pinned PPT Master commands for the document-artifact pipeline. It does not execute commands or overwrite a source deck. Preserve unselected slides and report `visual-unverified` when no renderer and visual review are available.
20
22
  - `aso_import` reads workspace-relative CSV/XLSX exports. Preserve nulls and source metadata. It is export intake, not a live Sensor Tower connection; do not call unregistered market-data tool names.