aiwg 2026.9.4 → 2026.9.6

This diff represents the content of publicly available package versions that have been released to one of the supported registries. The information contained in this diff is provided for informational purposes only and reflects changes between package versions as they appear in their respective public registries.
Files changed (748) hide show
  1. package/README.md +1414 -900
  2. package/THIRD_PARTY_NOTICES.md +16 -0
  3. package/agentic/code/addons/agent-loop/README.md +2 -0
  4. package/agentic/code/addons/agent-loop/docs/quickstart.md +18 -13
  5. package/agentic/code/addons/aiwg-dev/docs/overview.md +11 -2
  6. package/agentic/code/addons/aiwg-dev/docs/quickstart.md +8 -4
  7. package/agentic/code/addons/aiwg-utils/agents/aiwg-finder.md +4 -0
  8. package/agentic/code/addons/aiwg-utils/agents/session-analyst.md +44 -0
  9. package/agentic/code/addons/aiwg-utils/commands/sessions.md +33 -0
  10. package/agentic/code/addons/aiwg-utils/docs/overview.md +16 -6
  11. package/agentic/code/addons/aiwg-utils/flows/capabilities/session-evidence-collect.yaml +13 -0
  12. package/agentic/code/addons/aiwg-utils/flows/capabilities/session-evidence-synthesize.yaml +13 -0
  13. package/agentic/code/addons/aiwg-utils/flows/session-investigation.yaml +17 -0
  14. package/agentic/code/addons/aiwg-utils/manifest.json +16 -4
  15. package/agentic/code/addons/aiwg-utils/skills/aiwg-language-map/SKILL.md +10 -0
  16. package/agentic/code/addons/aiwg-utils/skills/cost-history/SKILL.md +5 -0
  17. package/agentic/code/addons/aiwg-utils/skills/session/SKILL.md +3 -0
  18. package/agentic/code/addons/aiwg-utils/skills/session-explore/SKILL.md +105 -0
  19. package/agentic/code/addons/aiwg-utils/skills/session-explore/references/recipes.md +146 -0
  20. package/agentic/code/addons/aiwg-utils/skills/session-harvest/SKILL.md +73 -0
  21. package/agentic/code/addons/aiwg-utils/skills/summarize-transcript/SKILL.md +9 -0
  22. package/agentic/code/addons/auto-memory/docs/overview.md +23 -4
  23. package/agentic/code/addons/civic-action/docs/overview.md +8 -1
  24. package/agentic/code/addons/civic-action/docs/quickstart.md +3 -1
  25. package/agentic/code/addons/compound-memory/docs/overview.md +12 -2
  26. package/agentic/code/addons/daemon/docs/quickstart.md +3 -1
  27. package/agentic/code/addons/dataset-intelligence/docs/overview.md +10 -3
  28. package/agentic/code/addons/dataset-intelligence/skills/dataset-intake/SKILL.md +6 -0
  29. package/agentic/code/addons/guided-implementation/docs/overview.md +22 -3
  30. package/agentic/code/addons/line-memory/docs/overview.md +7 -4
  31. package/agentic/code/addons/network-analysis/CHANGELOG.md +18 -0
  32. package/agentic/code/addons/network-analysis/README.md +108 -0
  33. package/agentic/code/addons/network-analysis/docs/integrations.md +20 -0
  34. package/agentic/code/addons/network-analysis/docs/maintainer-guide.md +105 -0
  35. package/agentic/code/addons/network-analysis/docs/offline-analysis.md +80 -0
  36. package/agentic/code/addons/network-analysis/docs/operator-guide.md +119 -0
  37. package/agentic/code/addons/network-analysis/docs/overview.md +30 -0
  38. package/agentic/code/addons/network-analysis/docs/release-checklist.md +48 -0
  39. package/agentic/code/addons/network-analysis/docs/termshark-handoff.md +72 -0
  40. package/agentic/code/addons/network-analysis/manifest.json +78 -0
  41. package/agentic/code/addons/network-analysis/recipes/README.md +42 -0
  42. package/agentic/code/addons/network-analysis/recipes/beaconing-timing.json +15 -0
  43. package/agentic/code/addons/network-analysis/recipes/before-after.json +15 -0
  44. package/agentic/code/addons/network-analysis/recipes/dns.json +15 -0
  45. package/agentic/code/addons/network-analysis/recipes/endpoints-conversations.json +15 -0
  46. package/agentic/code/addons/network-analysis/recipes/http-metadata.json +15 -0
  47. package/agentic/code/addons/network-analysis/recipes/overview.json +27 -0
  48. package/agentic/code/addons/network-analysis/recipes/stream-selection.json +15 -0
  49. package/agentic/code/addons/network-analysis/recipes/tcp-health.json +15 -0
  50. package/agentic/code/addons/network-analysis/recipes/tls.json +15 -0
  51. package/agentic/code/addons/network-analysis/rules/RULES-INDEX.md +5 -0
  52. package/agentic/code/addons/network-analysis/rules/network-analysis-safety.md +29 -0
  53. package/agentic/code/addons/network-analysis/schemas/network-analysis-contracts.md +19 -0
  54. package/agentic/code/addons/network-analysis/skills/analyze-network-capture/SKILL.md +65 -0
  55. package/agentic/code/addons/network-analysis/templates/analysis-request.md +18 -0
  56. package/agentic/code/addons/network-analysis/templates/termshark-handoff.md +19 -0
  57. package/agentic/code/addons/prose-integration/docs/overview.md +16 -4
  58. package/agentic/code/addons/semantic-memory/skills/memory-ingest/SKILL.md +9 -0
  59. package/agentic/code/addons/testing-quality/README.md +97 -112
  60. package/agentic/code/addons/testing-quality/adapters/README.md +26 -0
  61. package/agentic/code/addons/testing-quality/adapters/XML-FORMATS.md +43 -0
  62. package/agentic/code/addons/testing-quality/adapters/pytest_reporter.py +102 -0
  63. package/agentic/code/addons/testing-quality/agents/conformance-steward.md +32 -0
  64. package/agentic/code/addons/testing-quality/agents/test-normalization-engineer.md +36 -0
  65. package/agentic/code/addons/testing-quality/agents/test-oracle-reviewer.md +37 -0
  66. package/agentic/code/addons/testing-quality/commands/test-conformance.mjs +136 -0
  67. package/agentic/code/addons/testing-quality/docs/conformance-workflow.md +223 -0
  68. package/agentic/code/addons/testing-quality/docs/overview.md +38 -14
  69. package/agentic/code/addons/testing-quality/docs/quickstart.md +51 -33
  70. package/agentic/code/addons/testing-quality/flows/capabilities/conformance-apply.yaml +29 -0
  71. package/agentic/code/addons/testing-quality/flows/capabilities/conformance-assess.yaml +43 -0
  72. package/agentic/code/addons/testing-quality/flows/capabilities/conformance-collect.yaml +39 -0
  73. package/agentic/code/addons/testing-quality/flows/capabilities/conformance-inventory.yaml +27 -0
  74. package/agentic/code/addons/testing-quality/flows/capabilities/conformance-plan.yaml +34 -0
  75. package/agentic/code/addons/testing-quality/flows/capabilities/conformance-platform-research.yaml +29 -0
  76. package/agentic/code/addons/testing-quality/flows/capabilities/conformance-protocol.yaml +24 -0
  77. package/agentic/code/addons/testing-quality/flows/capabilities/conformance-report.yaml +24 -0
  78. package/agentic/code/addons/testing-quality/flows/capabilities/conformance-review.yaml +38 -0
  79. package/agentic/code/addons/testing-quality/flows/capabilities/conformance-verify.yaml +42 -0
  80. package/agentic/code/addons/testing-quality/flows/test-conformance-audit.yaml +69 -0
  81. package/agentic/code/addons/testing-quality/flows/test-conformance-normalize.yaml +55 -0
  82. package/agentic/code/addons/testing-quality/flows/test-platform-qualification.yaml +18 -0
  83. package/agentic/code/addons/testing-quality/lib/assessment.mjs +213 -0
  84. package/agentic/code/addons/testing-quality/lib/collector.mjs +194 -0
  85. package/agentic/code/addons/testing-quality/lib/contracts.mjs +90 -0
  86. package/agentic/code/addons/testing-quality/lib/controls.mjs +167 -0
  87. package/agentic/code/addons/testing-quality/lib/coverage.mjs +63 -0
  88. package/agentic/code/addons/testing-quality/lib/inventory.mjs +107 -0
  89. package/agentic/code/addons/testing-quality/lib/normalization.mjs +156 -0
  90. package/agentic/code/addons/testing-quality/lib/profiles.mjs +104 -0
  91. package/agentic/code/addons/testing-quality/lib/research.mjs +102 -0
  92. package/agentic/code/addons/testing-quality/lib/results.mjs +228 -0
  93. package/agentic/code/addons/testing-quality/lib/templates.mjs +124 -0
  94. package/agentic/code/addons/testing-quality/lib/workspace.mjs +74 -0
  95. package/agentic/code/addons/testing-quality/lib/xml-results.mjs +131 -0
  96. package/agentic/code/addons/testing-quality/manifest.json +93 -3
  97. package/agentic/code/addons/testing-quality/profiles/dotnet-vstest.json +151 -0
  98. package/agentic/code/addons/testing-quality/profiles/generic.json +141 -0
  99. package/agentic/code/addons/testing-quality/profiles/go.json +144 -0
  100. package/agentic/code/addons/testing-quality/profiles/java-junit.json +143 -0
  101. package/agentic/code/addons/testing-quality/profiles/javascript-jest.json +150 -0
  102. package/agentic/code/addons/testing-quality/profiles/javascript-node.json +146 -0
  103. package/agentic/code/addons/testing-quality/profiles/javascript-vitest.json +165 -0
  104. package/agentic/code/addons/testing-quality/profiles/python-pytest.json +177 -0
  105. package/agentic/code/addons/testing-quality/profiles/rust-cargo.json +145 -0
  106. package/agentic/code/addons/testing-quality/research/primary-sources.json +270 -0
  107. package/agentic/code/addons/testing-quality/research/principles.json +41 -0
  108. package/agentic/code/addons/testing-quality/research/tool-recommendations.json +1085 -0
  109. package/agentic/code/addons/testing-quality/rules/test-conformance-evidence.md +28 -0
  110. package/agentic/code/addons/testing-quality/schemas/conformance-protocol.v1.schema.json +620 -0
  111. package/agentic/code/addons/testing-quality/schemas/custom-template.v1.schema.json +134 -0
  112. package/agentic/code/addons/testing-quality/schemas/negative-control-receipt.v1.schema.json +201 -0
  113. package/agentic/code/addons/testing-quality/schemas/normalization-plan.v1.schema.json +180 -0
  114. package/agentic/code/addons/testing-quality/schemas/normalization-receipt.v1.schema.json +294 -0
  115. package/agentic/code/addons/testing-quality/schemas/normalized-results.v1.schema.json +141 -0
  116. package/agentic/code/addons/testing-quality/schemas/test-conformance-assessment.v1.schema.json +229 -0
  117. package/agentic/code/addons/testing-quality/schemas/test-conformance-research.v1.schema.json +527 -0
  118. package/agentic/code/addons/testing-quality/schemas/test-coverage.v1.schema.json +325 -0
  119. package/agentic/code/addons/testing-quality/schemas/test-inventory.v1.schema.json +272 -0
  120. package/agentic/code/addons/testing-quality/schemas/test-review.v1.schema.json +208 -0
  121. package/agentic/code/addons/testing-quality/schemas/test-run-receipt.v1.schema.json +949 -0
  122. package/agentic/code/addons/testing-quality/schemas/test-sample.v1.schema.json +261 -0
  123. package/agentic/code/addons/testing-quality/skills/test-conformance/SKILL.md +36 -0
  124. package/agentic/code/addons/testing-quality/skills/test-normalize/SKILL.md +37 -0
  125. package/agentic/code/addons/testing-quality/skills/test-platform-research/SKILL.md +37 -0
  126. package/agentic/code/addons/testing-quality/templates/adapter-qualification.md +26 -0
  127. package/agentic/code/addons/testing-quality/templates/conformance-report.md +44 -0
  128. package/agentic/code/addons/testing-quality/templates/normalization-plan.md +30 -0
  129. package/agentic/code/addons/testing-quality/templates/platforms/dotnet-vstest/BoundaryExampleTests.cs +8 -0
  130. package/agentic/code/addons/testing-quality/templates/platforms/dotnet-vstest/RESEARCH.md +11 -0
  131. package/agentic/code/addons/testing-quality/templates/platforms/dotnet-vstest/protocol.example.json +126 -0
  132. package/agentic/code/addons/testing-quality/templates/platforms/dotnet-vstest/test-review.example.md +9 -0
  133. package/agentic/code/addons/testing-quality/templates/platforms/generic/RESEARCH.md +10 -0
  134. package/agentic/code/addons/testing-quality/templates/platforms/generic/oracle.example.test.mjs +4 -0
  135. package/agentic/code/addons/testing-quality/templates/platforms/generic/protocol.example.json +116 -0
  136. package/agentic/code/addons/testing-quality/templates/platforms/generic/test-review.example.md +9 -0
  137. package/agentic/code/addons/testing-quality/templates/platforms/go/RESEARCH.md +10 -0
  138. package/agentic/code/addons/testing-quality/templates/platforms/go/oracle_example_test.go +5 -0
  139. package/agentic/code/addons/testing-quality/templates/platforms/go/protocol.example.json +119 -0
  140. package/agentic/code/addons/testing-quality/templates/platforms/go/test-review.example.md +9 -0
  141. package/agentic/code/addons/testing-quality/templates/platforms/java-junit/BoundaryExampleTest.java +7 -0
  142. package/agentic/code/addons/testing-quality/templates/platforms/java-junit/RESEARCH.md +11 -0
  143. package/agentic/code/addons/testing-quality/templates/platforms/java-junit/protocol.example.json +118 -0
  144. package/agentic/code/addons/testing-quality/templates/platforms/java-junit/test-review.example.md +9 -0
  145. package/agentic/code/addons/testing-quality/templates/platforms/javascript-jest/RESEARCH.md +10 -0
  146. package/agentic/code/addons/testing-quality/templates/platforms/javascript-jest/jest.config.example.mjs +2 -0
  147. package/agentic/code/addons/testing-quality/templates/platforms/javascript-jest/protocol.example.json +125 -0
  148. package/agentic/code/addons/testing-quality/templates/platforms/javascript-jest/test-review.example.md +9 -0
  149. package/agentic/code/addons/testing-quality/templates/platforms/javascript-node/RESEARCH.md +10 -0
  150. package/agentic/code/addons/testing-quality/templates/platforms/javascript-node/oracle.example.test.mjs +5 -0
  151. package/agentic/code/addons/testing-quality/templates/platforms/javascript-node/protocol.example.json +121 -0
  152. package/agentic/code/addons/testing-quality/templates/platforms/javascript-node/test-review.example.md +9 -0
  153. package/agentic/code/addons/testing-quality/templates/platforms/javascript-vitest/RESEARCH.md +11 -0
  154. package/agentic/code/addons/testing-quality/templates/platforms/javascript-vitest/protocol.example.json +140 -0
  155. package/agentic/code/addons/testing-quality/templates/platforms/javascript-vitest/test-review.example.md +9 -0
  156. package/agentic/code/addons/testing-quality/templates/platforms/javascript-vitest/vitest.config.example.mjs +2 -0
  157. package/agentic/code/addons/testing-quality/templates/platforms/python-pytest/RESEARCH.md +10 -0
  158. package/agentic/code/addons/testing-quality/templates/platforms/python-pytest/protocol.example.json +147 -0
  159. package/agentic/code/addons/testing-quality/templates/platforms/python-pytest/pytest.example.ini +4 -0
  160. package/agentic/code/addons/testing-quality/templates/platforms/python-pytest/test-review.example.md +9 -0
  161. package/agentic/code/addons/testing-quality/templates/platforms/rust-cargo/RESEARCH.md +11 -0
  162. package/agentic/code/addons/testing-quality/templates/platforms/rust-cargo/oracle_example.rs +8 -0
  163. package/agentic/code/addons/testing-quality/templates/platforms/rust-cargo/protocol.example.json +120 -0
  164. package/agentic/code/addons/testing-quality/templates/platforms/rust-cargo/test-review.example.md +9 -0
  165. package/agentic/code/addons/testing-quality/templates/protocol-review.md +38 -0
  166. package/agentic/code/addons/testing-quality/templates/test-review.md +26 -0
  167. package/agentic/code/addons/testing-quality/templates/tool-recommendation.md +21 -0
  168. package/agentic/code/addons/voice-framework/README.md +16 -1
  169. package/agentic/code/addons/voice-framework/docs/contextual-diagnostics-evaluation.md +52 -0
  170. package/agentic/code/addons/voice-framework/docs/contextual-diagnostics.md +66 -0
  171. package/agentic/code/addons/voice-framework/docs/natural-voice/ADR-001-evidence-and-ownership.md +29 -0
  172. package/agentic/code/addons/voice-framework/docs/natural-voice/evidence-ledger.v1.json +434 -0
  173. package/agentic/code/addons/voice-framework/docs/natural-voice/validate-ledger.mjs +52 -0
  174. package/agentic/code/addons/voice-framework/docs/natural-voice/verification-2026-09-06.md +19 -0
  175. package/agentic/code/addons/voice-framework/docs/overview.md +32 -7
  176. package/agentic/code/addons/voice-framework/docs/quickstart.md +25 -9
  177. package/agentic/code/addons/voice-framework/docs/voice-output-impact.md +54 -0
  178. package/agentic/code/addons/voice-framework/docs/writing-workflows.md +53 -0
  179. package/agentic/code/addons/voice-framework/flows/voice-critique-correction.flow.yaml +595 -0
  180. package/agentic/code/addons/voice-framework/manifest.json +8 -4
  181. package/agentic/code/addons/voice-framework/skills/manifest.json +53 -12
  182. package/agentic/code/addons/voice-framework/skills/output-mode-guide/SKILL.md +12 -0
  183. package/agentic/code/addons/voice-framework/skills/voice-analyze/SKILL.md +18 -0
  184. package/agentic/code/addons/voice-framework/skills/voice-apply/SKILL.md +47 -7
  185. package/agentic/code/addons/voice-framework/skills/voice-blend/SKILL.md +8 -0
  186. package/agentic/code/addons/voice-framework/skills/voice-create/SKILL.md +19 -0
  187. package/agentic/code/addons/writing-quality/README.md +8 -2
  188. package/agentic/code/addons/writing-quality/agents/content-diversifier.md +31 -28
  189. package/agentic/code/addons/writing-quality/agents/prompt-optimizer.md +25 -267
  190. package/agentic/code/addons/writing-quality/agents/writing-validator.md +28 -34
  191. package/agentic/code/addons/writing-quality/context/quick-reference.md +4 -4
  192. package/agentic/code/addons/writing-quality/core/philosophy.md +6 -6
  193. package/agentic/code/addons/writing-quality/manifest.json +1 -1
  194. package/agentic/code/addons/writing-quality/skills/ai-pattern-detection/SKILL.md +20 -16
  195. package/agentic/code/addons/writing-quality/skills/ai-pattern-detection/references/quick-patterns.md +6 -4
  196. package/agentic/code/addons/writing-quality/skills/ai-pattern-detection/scripts/pattern_scanner.py +13 -5
  197. package/agentic/code/addons/writing-quality/validation/banned-patterns.md +5 -7
  198. package/agentic/code/addons/writing-quality/validation/scoring-config.json +9 -9
  199. package/agentic/code/addons/writing-quality/validation/validation-checklist.md +22 -26
  200. package/agentic/code/frameworks/forensics-complete/README.md +4 -0
  201. package/agentic/code/frameworks/forensics-complete/agents/network-analyst.md +6 -0
  202. package/agentic/code/frameworks/forensics-complete/docs/packet-evidence-integration.md +78 -0
  203. package/agentic/code/frameworks/forensics-complete/templates/chain-of-custody.md +13 -0
  204. package/agentic/code/frameworks/knowledge-base/docs/overview.md +20 -2
  205. package/agentic/code/frameworks/media-curator/docs/overview.md +17 -6
  206. package/agentic/code/frameworks/media-marketing-kit/agents/content-writer.md +16 -7
  207. package/agentic/code/frameworks/media-marketing-kit/agents/copywriter.md +9 -0
  208. package/agentic/code/frameworks/media-marketing-kit/agents/email-marketer.md +9 -0
  209. package/agentic/code/frameworks/media-marketing-kit/agents/social-media-specialist.md +9 -0
  210. package/agentic/code/frameworks/media-marketing-kit/docs/overview.md +23 -15
  211. package/agentic/code/frameworks/media-marketing-kit/docs/quickstart.md +28 -8
  212. package/agentic/code/frameworks/media-marketing-kit/skills/pr-launch/SKILL.md +9 -0
  213. package/agentic/code/frameworks/media-marketing-kit/templates/content/blog-post-template.md +5 -5
  214. package/agentic/code/frameworks/media-marketing-kit/templates/social/social-post-template.md +6 -0
  215. package/agentic/code/frameworks/ops-complete/README.md +1 -0
  216. package/agentic/code/frameworks/ops-complete/docs/overview.md +17 -7
  217. package/agentic/code/frameworks/ops-complete/docs/packet-verification.md +53 -0
  218. package/agentic/code/frameworks/ops-complete/docs/quickstart.md +23 -3
  219. package/agentic/code/frameworks/research-complete/README.md +1 -0
  220. package/agentic/code/frameworks/research-complete/config/source-types.yaml +11 -0
  221. package/agentic/code/frameworks/research-complete/docs/overview.md +39 -26
  222. package/agentic/code/frameworks/research-complete/docs/packet-evidence.md +62 -0
  223. package/agentic/code/frameworks/research-complete/docs/quickstart.md +30 -12
  224. package/agentic/code/frameworks/research-complete/skills/induct-research/SKILL.md +2 -0
  225. package/agentic/code/frameworks/research-complete/skills/research-acquire/SKILL.md +2 -0
  226. package/agentic/code/frameworks/research-complete/skills/research-workflow/SKILL.md +2 -0
  227. package/agentic/code/frameworks/research-complete/skills/source-types/SKILL.md +1 -1
  228. package/agentic/code/frameworks/research-complete/templates/manifest.json +14 -3
  229. package/agentic/code/frameworks/research-complete/templates/reference-packet-evidence.md +75 -0
  230. package/agentic/code/frameworks/sdlc-complete/schemas/verification-contract.schema.json +1 -0
  231. package/agentic/code/frameworks/sdlc-complete/skills/issue-planner/SKILL.md +19 -0
  232. package/agentic/code/frameworks/sdlc-complete/templates/deepseek-harness/AGENTS.md.aiwg-template +20 -0
  233. package/agentic/code/frameworks/sdlc-complete/templates/test/manifest.json +2 -0
  234. package/agentic/code/frameworks/sdlc-complete/templates/test/packet-evidence-defect.md +19 -0
  235. package/agentic/code/frameworks/sdlc-complete/templates/test/packet-evidence-test-plan.md +27 -0
  236. package/agentic/code/frameworks/security-engineering/docs/network-control-review.md +21 -0
  237. package/agentic/code/plugins/agent-loop/.claude-plugin/plugin.json +1 -1
  238. package/agentic/code/plugins/agent-loop/README.md +2 -0
  239. package/agentic/code/plugins/agent-loop/docs/quickstart.md +18 -13
  240. package/agentic/code/plugins/agent-persistence/.claude-plugin/plugin.json +1 -1
  241. package/agentic/code/plugins/agentic-installer/.claude-plugin/plugin.json +1 -1
  242. package/agentic/code/plugins/aiwg-dev/.claude-plugin/plugin.json +1 -1
  243. package/agentic/code/plugins/aiwg-dev/docs/overview.md +11 -2
  244. package/agentic/code/plugins/aiwg-dev/docs/quickstart.md +8 -4
  245. package/agentic/code/plugins/aiwg-evals/.claude-plugin/plugin.json +1 -1
  246. package/agentic/code/plugins/auto-memory/.claude-plugin/plugin.json +1 -1
  247. package/agentic/code/plugins/auto-memory/docs/overview.md +23 -4
  248. package/agentic/code/plugins/browser-control/.claude-plugin/plugin.json +1 -1
  249. package/agentic/code/plugins/color-palette/.claude-plugin/plugin.json +1 -1
  250. package/agentic/code/plugins/compound-memory/.claude-plugin/plugin.json +1 -1
  251. package/agentic/code/plugins/compound-memory/docs/overview.md +12 -2
  252. package/agentic/code/plugins/context-curator/.claude-plugin/plugin.json +1 -1
  253. package/agentic/code/plugins/daemon/.claude-plugin/plugin.json +1 -1
  254. package/agentic/code/plugins/daemon/docs/quickstart.md +3 -1
  255. package/agentic/code/plugins/doc-intelligence/.claude-plugin/plugin.json +1 -1
  256. package/agentic/code/plugins/droid-bridge/.claude-plugin/plugin.json +1 -1
  257. package/agentic/code/plugins/forensics/.claude-plugin/plugin.json +1 -1
  258. package/agentic/code/plugins/forensics/agents/network-analyst.md +6 -0
  259. package/agentic/code/plugins/guided-implementation/.claude-plugin/plugin.json +1 -1
  260. package/agentic/code/plugins/guided-implementation/docs/overview.md +22 -3
  261. package/agentic/code/plugins/hooks/.claude-plugin/plugin.json +1 -1
  262. package/agentic/code/plugins/knowledge-base/.claude-plugin/plugin.json +1 -1
  263. package/agentic/code/plugins/line-memory/.claude-plugin/plugin.json +1 -1
  264. package/agentic/code/plugins/line-memory/docs/overview.md +7 -4
  265. package/agentic/code/plugins/llm-wiki/.claude-plugin/plugin.json +1 -1
  266. package/agentic/code/plugins/marketing/.claude-plugin/plugin.json +1 -1
  267. package/agentic/code/plugins/marketing/agents/content-writer.md +16 -7
  268. package/agentic/code/plugins/marketing/agents/copywriter.md +9 -0
  269. package/agentic/code/plugins/marketing/agents/email-marketer.md +9 -0
  270. package/agentic/code/plugins/marketing/agents/social-media-specialist.md +9 -0
  271. package/agentic/code/plugins/marketing/skills/pr-launch/SKILL.md +9 -0
  272. package/agentic/code/plugins/media-curator/.claude-plugin/plugin.json +1 -1
  273. package/agentic/code/plugins/nlp-prod/.claude-plugin/plugin.json +1 -1
  274. package/agentic/code/plugins/ops/.claude-plugin/plugin.json +1 -1
  275. package/agentic/code/plugins/prose-integration/.claude-plugin/plugin.json +1 -1
  276. package/agentic/code/plugins/prose-integration/docs/overview.md +16 -4
  277. package/agentic/code/plugins/research/.claude-plugin/plugin.json +1 -1
  278. package/agentic/code/plugins/research/skills/induct-research/SKILL.md +2 -0
  279. package/agentic/code/plugins/research/skills/research-acquire/SKILL.md +2 -0
  280. package/agentic/code/plugins/research/skills/research-workflow/SKILL.md +2 -0
  281. package/agentic/code/plugins/research/skills/source-types/SKILL.md +1 -1
  282. package/agentic/code/plugins/rlm/.claude-plugin/plugin.json +1 -1
  283. package/agentic/code/plugins/sdlc/.claude-plugin/plugin.json +1 -1
  284. package/agentic/code/plugins/sdlc/skills/issue-planner/SKILL.md +19 -0
  285. package/agentic/code/plugins/security-engineering/.claude-plugin/plugin.json +1 -1
  286. package/agentic/code/plugins/semantic-memory/.claude-plugin/plugin.json +1 -1
  287. package/agentic/code/plugins/semantic-memory/skills/memory-ingest/SKILL.md +9 -0
  288. package/agentic/code/plugins/skill-factory/.claude-plugin/plugin.json +1 -1
  289. package/agentic/code/plugins/star-prompt/.claude-plugin/plugin.json +1 -1
  290. package/agentic/code/plugins/testing-quality/.claude-plugin/plugin.json +2 -2
  291. package/agentic/code/plugins/testing-quality/README.md +97 -112
  292. package/agentic/code/plugins/testing-quality/adapters/README.md +26 -0
  293. package/agentic/code/plugins/testing-quality/adapters/XML-FORMATS.md +43 -0
  294. package/agentic/code/plugins/testing-quality/adapters/pytest_reporter.py +102 -0
  295. package/agentic/code/plugins/testing-quality/agents/conformance-steward.md +32 -0
  296. package/agentic/code/plugins/testing-quality/agents/test-normalization-engineer.md +36 -0
  297. package/agentic/code/plugins/testing-quality/agents/test-oracle-reviewer.md +37 -0
  298. package/agentic/code/plugins/testing-quality/commands/test-conformance.mjs +136 -0
  299. package/agentic/code/plugins/testing-quality/docs/conformance-workflow.md +223 -0
  300. package/agentic/code/plugins/testing-quality/docs/overview.md +38 -14
  301. package/agentic/code/plugins/testing-quality/docs/quickstart.md +51 -33
  302. package/agentic/code/plugins/testing-quality/flows/capabilities/conformance-apply.yaml +29 -0
  303. package/agentic/code/plugins/testing-quality/flows/capabilities/conformance-assess.yaml +43 -0
  304. package/agentic/code/plugins/testing-quality/flows/capabilities/conformance-collect.yaml +39 -0
  305. package/agentic/code/plugins/testing-quality/flows/capabilities/conformance-inventory.yaml +27 -0
  306. package/agentic/code/plugins/testing-quality/flows/capabilities/conformance-plan.yaml +34 -0
  307. package/agentic/code/plugins/testing-quality/flows/capabilities/conformance-platform-research.yaml +29 -0
  308. package/agentic/code/plugins/testing-quality/flows/capabilities/conformance-protocol.yaml +24 -0
  309. package/agentic/code/plugins/testing-quality/flows/capabilities/conformance-report.yaml +24 -0
  310. package/agentic/code/plugins/testing-quality/flows/capabilities/conformance-review.yaml +38 -0
  311. package/agentic/code/plugins/testing-quality/flows/capabilities/conformance-verify.yaml +42 -0
  312. package/agentic/code/plugins/testing-quality/flows/test-conformance-audit.yaml +69 -0
  313. package/agentic/code/plugins/testing-quality/flows/test-conformance-normalize.yaml +55 -0
  314. package/agentic/code/plugins/testing-quality/flows/test-platform-qualification.yaml +18 -0
  315. package/agentic/code/plugins/testing-quality/lib/assessment.mjs +213 -0
  316. package/agentic/code/plugins/testing-quality/lib/collector.mjs +194 -0
  317. package/agentic/code/plugins/testing-quality/lib/contracts.mjs +90 -0
  318. package/agentic/code/plugins/testing-quality/lib/controls.mjs +167 -0
  319. package/agentic/code/plugins/testing-quality/lib/coverage.mjs +63 -0
  320. package/agentic/code/plugins/testing-quality/lib/inventory.mjs +107 -0
  321. package/agentic/code/plugins/testing-quality/lib/normalization.mjs +156 -0
  322. package/agentic/code/plugins/testing-quality/lib/profiles.mjs +104 -0
  323. package/agentic/code/plugins/testing-quality/lib/research.mjs +102 -0
  324. package/agentic/code/plugins/testing-quality/lib/results.mjs +228 -0
  325. package/agentic/code/plugins/testing-quality/lib/templates.mjs +124 -0
  326. package/agentic/code/plugins/testing-quality/lib/workspace.mjs +74 -0
  327. package/agentic/code/plugins/testing-quality/lib/xml-results.mjs +131 -0
  328. package/agentic/code/plugins/testing-quality/manifest.json +93 -3
  329. package/agentic/code/plugins/testing-quality/profiles/dotnet-vstest.json +151 -0
  330. package/agentic/code/plugins/testing-quality/profiles/generic.json +141 -0
  331. package/agentic/code/plugins/testing-quality/profiles/go.json +144 -0
  332. package/agentic/code/plugins/testing-quality/profiles/java-junit.json +143 -0
  333. package/agentic/code/plugins/testing-quality/profiles/javascript-jest.json +150 -0
  334. package/agentic/code/plugins/testing-quality/profiles/javascript-node.json +146 -0
  335. package/agentic/code/plugins/testing-quality/profiles/javascript-vitest.json +165 -0
  336. package/agentic/code/plugins/testing-quality/profiles/python-pytest.json +177 -0
  337. package/agentic/code/plugins/testing-quality/profiles/rust-cargo.json +145 -0
  338. package/agentic/code/plugins/testing-quality/research/primary-sources.json +270 -0
  339. package/agentic/code/plugins/testing-quality/research/principles.json +41 -0
  340. package/agentic/code/plugins/testing-quality/research/tool-recommendations.json +1085 -0
  341. package/agentic/code/plugins/testing-quality/rules/test-conformance-evidence.md +28 -0
  342. package/agentic/code/plugins/testing-quality/schemas/conformance-protocol.v1.schema.json +620 -0
  343. package/agentic/code/plugins/testing-quality/schemas/custom-template.v1.schema.json +134 -0
  344. package/agentic/code/plugins/testing-quality/schemas/negative-control-receipt.v1.schema.json +201 -0
  345. package/agentic/code/plugins/testing-quality/schemas/normalization-plan.v1.schema.json +180 -0
  346. package/agentic/code/plugins/testing-quality/schemas/normalization-receipt.v1.schema.json +294 -0
  347. package/agentic/code/plugins/testing-quality/schemas/normalized-results.v1.schema.json +141 -0
  348. package/agentic/code/plugins/testing-quality/schemas/test-conformance-assessment.v1.schema.json +229 -0
  349. package/agentic/code/plugins/testing-quality/schemas/test-conformance-research.v1.schema.json +527 -0
  350. package/agentic/code/plugins/testing-quality/schemas/test-coverage.v1.schema.json +325 -0
  351. package/agentic/code/plugins/testing-quality/schemas/test-inventory.v1.schema.json +272 -0
  352. package/agentic/code/plugins/testing-quality/schemas/test-review.v1.schema.json +208 -0
  353. package/agentic/code/plugins/testing-quality/schemas/test-run-receipt.v1.schema.json +949 -0
  354. package/agentic/code/plugins/testing-quality/schemas/test-sample.v1.schema.json +261 -0
  355. package/agentic/code/plugins/testing-quality/skills/test-conformance/SKILL.md +36 -0
  356. package/agentic/code/plugins/testing-quality/skills/test-normalize/SKILL.md +37 -0
  357. package/agentic/code/plugins/testing-quality/skills/test-platform-research/SKILL.md +37 -0
  358. package/agentic/code/plugins/testing-quality/templates/adapter-qualification.md +26 -0
  359. package/agentic/code/plugins/testing-quality/templates/conformance-report.md +44 -0
  360. package/agentic/code/plugins/testing-quality/templates/normalization-plan.md +30 -0
  361. package/agentic/code/plugins/testing-quality/templates/platforms/dotnet-vstest/BoundaryExampleTests.cs +8 -0
  362. package/agentic/code/plugins/testing-quality/templates/platforms/dotnet-vstest/RESEARCH.md +11 -0
  363. package/agentic/code/plugins/testing-quality/templates/platforms/dotnet-vstest/protocol.example.json +126 -0
  364. package/agentic/code/plugins/testing-quality/templates/platforms/dotnet-vstest/test-review.example.md +9 -0
  365. package/agentic/code/plugins/testing-quality/templates/platforms/generic/RESEARCH.md +10 -0
  366. package/agentic/code/plugins/testing-quality/templates/platforms/generic/oracle.example.test.mjs +4 -0
  367. package/agentic/code/plugins/testing-quality/templates/platforms/generic/protocol.example.json +116 -0
  368. package/agentic/code/plugins/testing-quality/templates/platforms/generic/test-review.example.md +9 -0
  369. package/agentic/code/plugins/testing-quality/templates/platforms/go/RESEARCH.md +10 -0
  370. package/agentic/code/plugins/testing-quality/templates/platforms/go/oracle_example_test.go +5 -0
  371. package/agentic/code/plugins/testing-quality/templates/platforms/go/protocol.example.json +119 -0
  372. package/agentic/code/plugins/testing-quality/templates/platforms/go/test-review.example.md +9 -0
  373. package/agentic/code/plugins/testing-quality/templates/platforms/java-junit/BoundaryExampleTest.java +7 -0
  374. package/agentic/code/plugins/testing-quality/templates/platforms/java-junit/RESEARCH.md +11 -0
  375. package/agentic/code/plugins/testing-quality/templates/platforms/java-junit/protocol.example.json +118 -0
  376. package/agentic/code/plugins/testing-quality/templates/platforms/java-junit/test-review.example.md +9 -0
  377. package/agentic/code/plugins/testing-quality/templates/platforms/javascript-jest/RESEARCH.md +10 -0
  378. package/agentic/code/plugins/testing-quality/templates/platforms/javascript-jest/jest.config.example.mjs +2 -0
  379. package/agentic/code/plugins/testing-quality/templates/platforms/javascript-jest/protocol.example.json +125 -0
  380. package/agentic/code/plugins/testing-quality/templates/platforms/javascript-jest/test-review.example.md +9 -0
  381. package/agentic/code/plugins/testing-quality/templates/platforms/javascript-node/RESEARCH.md +10 -0
  382. package/agentic/code/plugins/testing-quality/templates/platforms/javascript-node/oracle.example.test.mjs +5 -0
  383. package/agentic/code/plugins/testing-quality/templates/platforms/javascript-node/protocol.example.json +121 -0
  384. package/agentic/code/plugins/testing-quality/templates/platforms/javascript-node/test-review.example.md +9 -0
  385. package/agentic/code/plugins/testing-quality/templates/platforms/javascript-vitest/RESEARCH.md +11 -0
  386. package/agentic/code/plugins/testing-quality/templates/platforms/javascript-vitest/protocol.example.json +140 -0
  387. package/agentic/code/plugins/testing-quality/templates/platforms/javascript-vitest/test-review.example.md +9 -0
  388. package/agentic/code/plugins/testing-quality/templates/platforms/javascript-vitest/vitest.config.example.mjs +2 -0
  389. package/agentic/code/plugins/testing-quality/templates/platforms/python-pytest/RESEARCH.md +10 -0
  390. package/agentic/code/plugins/testing-quality/templates/platforms/python-pytest/protocol.example.json +147 -0
  391. package/agentic/code/plugins/testing-quality/templates/platforms/python-pytest/pytest.example.ini +4 -0
  392. package/agentic/code/plugins/testing-quality/templates/platforms/python-pytest/test-review.example.md +9 -0
  393. package/agentic/code/plugins/testing-quality/templates/platforms/rust-cargo/RESEARCH.md +11 -0
  394. package/agentic/code/plugins/testing-quality/templates/platforms/rust-cargo/oracle_example.rs +8 -0
  395. package/agentic/code/plugins/testing-quality/templates/platforms/rust-cargo/protocol.example.json +120 -0
  396. package/agentic/code/plugins/testing-quality/templates/platforms/rust-cargo/test-review.example.md +9 -0
  397. package/agentic/code/plugins/testing-quality/templates/protocol-review.md +38 -0
  398. package/agentic/code/plugins/testing-quality/templates/test-review.md +26 -0
  399. package/agentic/code/plugins/testing-quality/templates/tool-recommendation.md +21 -0
  400. package/agentic/code/plugins/twelve-factor/.claude-plugin/plugin.json +1 -1
  401. package/agentic/code/plugins/uat-mcp/.claude-plugin/plugin.json +1 -1
  402. package/agentic/code/plugins/utils/.claude-plugin/plugin.json +1 -1
  403. package/agentic/code/plugins/utils/agents/aiwg-finder.md +4 -0
  404. package/agentic/code/plugins/utils/agents/session-analyst.md +44 -0
  405. package/agentic/code/plugins/utils/skills/aiwg-language-map/SKILL.md +10 -0
  406. package/agentic/code/plugins/utils/skills/cost-history/SKILL.md +5 -0
  407. package/agentic/code/plugins/utils/skills/session/SKILL.md +3 -0
  408. package/agentic/code/plugins/utils/skills/session-explore/SKILL.md +105 -0
  409. package/agentic/code/plugins/utils/skills/session-explore/references/recipes.md +146 -0
  410. package/agentic/code/plugins/utils/skills/session-harvest/SKILL.md +73 -0
  411. package/agentic/code/plugins/utils/skills/summarize-transcript/SKILL.md +9 -0
  412. package/agentic/code/plugins/validation-complete/.claude-plugin/plugin.json +1 -1
  413. package/agentic/code/plugins/verbalized-sampling/.claude-plugin/plugin.json +1 -1
  414. package/agentic/code/plugins/voice/.claude-plugin/plugin.json +1 -1
  415. package/agentic/code/plugins/voice/skills/manifest.json +53 -12
  416. package/agentic/code/plugins/voice/skills/output-mode-guide/SKILL.md +12 -0
  417. package/agentic/code/plugins/voice/skills/voice-analyze/SKILL.md +18 -0
  418. package/agentic/code/plugins/voice/skills/voice-apply/SKILL.md +47 -7
  419. package/agentic/code/plugins/voice/skills/voice-blend/SKILL.md +8 -0
  420. package/agentic/code/plugins/voice/skills/voice-create/SKILL.md +19 -0
  421. package/agentic/code/plugins/writing/.claude-plugin/plugin.json +1 -1
  422. package/agentic/code/plugins/writing/agents/content-diversifier.md +31 -28
  423. package/agentic/code/plugins/writing/agents/prompt-optimizer.md +25 -267
  424. package/agentic/code/plugins/writing/agents/writing-validator.md +28 -34
  425. package/agentic/code/plugins/writing/skills/ai-pattern-detection/SKILL.md +20 -16
  426. package/agentic/code/plugins/writing/skills/ai-pattern-detection/references/quick-patterns.md +6 -4
  427. package/agentic/code/plugins/writing/skills/ai-pattern-detection/scripts/pattern_scanner.py +13 -5
  428. package/agentic/code/providers/capability-matrix.yaml +45 -3
  429. package/agentic/code/providers/deepseek-harness/README.md +8 -0
  430. package/agentic/code/providers/deepseek-harness/aiwg.cordis.patch.yml +10 -0
  431. package/agentic/code/providers/model-capabilities.v1.json +11 -0
  432. package/agentic/code/providers/model-catalog.v1.json +8 -0
  433. package/bin/aiwg.mjs +1 -0
  434. package/dist/src/api/index.d.ts +14 -0
  435. package/dist/src/api/index.js +14 -0
  436. package/dist/src/artifacts/corpus-tools/source-types.js +1 -0
  437. package/dist/src/artifacts/index-files.js +17 -2
  438. package/dist/src/artifacts/repair.js +55 -6
  439. package/dist/src/catalog/cli.js +21 -7
  440. package/dist/src/catalog/cli.mjs +22 -7
  441. package/dist/src/cli/agent-spawn.js +10 -1
  442. package/dist/src/cli/handlers/artifacts.js +22 -3
  443. package/dist/src/cli/handlers/help.js +3 -0
  444. package/dist/src/cli/handlers/index.js +5 -1
  445. package/dist/src/cli/handlers/models.js +2 -2
  446. package/dist/src/cli/handlers/output-mode.js +1 -1
  447. package/dist/src/cli/handlers/runtime-info.js +1 -1
  448. package/dist/src/cli/handlers/sessions.js +55 -15
  449. package/dist/src/cli/handlers/steward.js +1 -1
  450. package/dist/src/cli/handlers/subcommands.js +5 -0
  451. package/dist/src/cli/handlers/writer-profile.js +110 -0
  452. package/dist/src/cli/handlers/writing.js +122 -0
  453. package/dist/src/cli/router.js +5 -1
  454. package/dist/src/config/project-artifacts-runtime.mjs +33 -1
  455. package/dist/src/config/project-artifacts.js +1 -1
  456. package/dist/src/dataset/fortemi-dataset-execution.d.ts +23 -0
  457. package/dist/src/dataset/fortemi-dataset-execution.js +158 -0
  458. package/dist/src/dataset/fortemi-live-qualification.d.ts +4 -2
  459. package/dist/src/dataset/fortemi-live-qualification.js +20 -21
  460. package/dist/src/dataset/fortemi-run-receipt.d.ts +44 -0
  461. package/dist/src/dataset/fortemi-run-receipt.js +74 -0
  462. package/dist/src/dataset/index.d.ts +2 -0
  463. package/dist/src/dataset/index.js +2 -0
  464. package/dist/src/extensions/commands/definitions.js +24 -0
  465. package/dist/src/extensions/manifest.js +3 -0
  466. package/dist/src/mcp/server.mjs +2 -0
  467. package/dist/src/mcp/tools/writer-profiles.mjs +40 -0
  468. package/dist/src/models/provider-policy.js +1 -1
  469. package/dist/src/network-analysis/analyzer.d.ts +80 -0
  470. package/dist/src/network-analysis/analyzer.js +667 -0
  471. package/dist/src/network-analysis/citations.d.ts +76 -0
  472. package/dist/src/network-analysis/citations.js +107 -0
  473. package/dist/src/network-analysis/forensics.d.ts +131 -0
  474. package/dist/src/network-analysis/forensics.js +132 -0
  475. package/dist/src/network-analysis/governance.d.ts +94 -0
  476. package/dist/src/network-analysis/governance.js +216 -0
  477. package/dist/src/network-analysis/index.d.ts +10 -0
  478. package/dist/src/network-analysis/index.js +10 -0
  479. package/dist/src/network-analysis/probe.d.ts +90 -0
  480. package/dist/src/network-analysis/probe.js +405 -0
  481. package/dist/src/network-analysis/recipes.d.ts +102 -0
  482. package/dist/src/network-analysis/recipes.js +88 -0
  483. package/dist/src/network-analysis/research.d.ts +119 -0
  484. package/dist/src/network-analysis/research.js +181 -0
  485. package/dist/src/network-analysis/termshark.d.ts +95 -0
  486. package/dist/src/network-analysis/termshark.js +252 -0
  487. package/dist/src/network-analysis/verification.d.ts +106 -0
  488. package/dist/src/network-analysis/verification.js +171 -0
  489. package/dist/src/output-modes/registry.js +37 -6
  490. package/dist/src/output-modes/runtime.js +164 -28
  491. package/dist/src/providers/provider-definitions.js +49 -0
  492. package/dist/src/providers/provider-inventory.js +1 -0
  493. package/dist/src/providers/transformation-receipt.js +3 -2
  494. package/dist/src/sessions/adapters/deepseek-harness.js +178 -0
  495. package/dist/src/sessions/batch-import.js +7 -0
  496. package/dist/src/sessions/contracts.js +2 -1
  497. package/dist/src/sessions/index.js +1 -0
  498. package/dist/src/sessions/workspace-discovery.js +5 -1
  499. package/dist/src/skills/deployer.js +6 -6
  500. package/dist/src/smiths/context-pipeline/workspace-context.js +7 -0
  501. package/dist/src/writing/channel-packs.d.ts +9 -0
  502. package/dist/src/writing/channel-packs.js +13 -0
  503. package/dist/src/writing/contextual-diagnostics.d.ts +66 -0
  504. package/dist/src/writing/contextual-diagnostics.js +142 -0
  505. package/dist/src/writing/example-generator.js +7 -6
  506. package/dist/src/writing/exemplar-selection.d.ts +164 -0
  507. package/dist/src/writing/exemplar-selection.js +186 -0
  508. package/dist/src/writing/fidelity.d.ts +22 -0
  509. package/dist/src/writing/fidelity.js +61 -0
  510. package/dist/src/writing/validation-engine.js +32 -15
  511. package/dist/src/writing/voice-evaluation.d.ts +547 -0
  512. package/dist/src/writing/voice-evaluation.js +301 -0
  513. package/dist/src/writing/voice-revision.d.ts +170 -0
  514. package/dist/src/writing/voice-revision.js +201 -0
  515. package/dist/src/writing/writer-migration.d.ts +124 -0
  516. package/dist/src/writing/writer-migration.js +216 -0
  517. package/dist/src/writing/writer-profile-legacy.d.ts +12 -0
  518. package/dist/src/writing/writer-profile-legacy.js +145 -0
  519. package/dist/src/writing/writer-profile-store.d.ts +24 -0
  520. package/dist/src/writing/writer-profile-store.js +117 -0
  521. package/dist/src/writing/writer-profile.d.ts +605 -0
  522. package/dist/src/writing/writer-profile.js +222 -0
  523. package/dist/src/writing/writing-brief.d.ts +423 -0
  524. package/dist/src/writing/writing-brief.js +166 -0
  525. package/dist/src/writing/writing-channels.d.ts +46 -0
  526. package/dist/src/writing/writing-channels.js +63 -0
  527. package/dist/src/writing/writing-consumer.d.ts +36 -0
  528. package/dist/src/writing/writing-consumer.js +39 -0
  529. package/dist/src/writing/writing-receipt.d.ts +1569 -0
  530. package/dist/src/writing/writing-receipt.js +266 -0
  531. package/docs/_manifest.json +210 -20
  532. package/docs/addons/agent-loop/quickstart.md +18 -13
  533. package/docs/addons/aiwg-dev/overview.md +11 -2
  534. package/docs/addons/aiwg-dev/quickstart.md +8 -4
  535. package/docs/addons/aiwg-utils/overview.md +16 -6
  536. package/docs/addons/auto-memory/overview.md +23 -4
  537. package/docs/addons/civic-action/overview.md +8 -1
  538. package/docs/addons/civic-action/quickstart.md +3 -1
  539. package/docs/addons/compound-memory/overview.md +12 -2
  540. package/docs/addons/daemon/quickstart.md +3 -1
  541. package/docs/addons/dataset-intelligence/overview.md +10 -3
  542. package/docs/addons/guided-implementation/overview.md +22 -3
  543. package/docs/addons/line-memory/overview.md +7 -4
  544. package/docs/addons/network-analysis/integrations.md +20 -0
  545. package/docs/addons/network-analysis/maintainer-guide.md +105 -0
  546. package/docs/addons/network-analysis/offline-analysis.md +80 -0
  547. package/docs/addons/network-analysis/operator-guide.md +119 -0
  548. package/docs/addons/network-analysis/overview.md +30 -0
  549. package/docs/addons/network-analysis/release-checklist.md +48 -0
  550. package/docs/addons/network-analysis/termshark-handoff.md +72 -0
  551. package/docs/addons/prose-integration/overview.md +16 -4
  552. package/docs/addons/ralph/quickstart.md +19 -13
  553. package/docs/addons/rlm/deployment-guide.md +2 -2
  554. package/docs/addons/rlm/multi-provider-guide.md +1 -1
  555. package/docs/addons/testing-quality/conformance-workflow.md +223 -0
  556. package/docs/addons/testing-quality/overview.md +38 -14
  557. package/docs/addons/testing-quality/quickstart.md +51 -33
  558. package/docs/addons/voice-framework/overview.md +32 -7
  559. package/docs/addons/voice-framework/quickstart.md +25 -9
  560. package/docs/architecture/adr-deepseek-harness-runtime.md +70 -0
  561. package/docs/architecture/network-analysis.md +147 -0
  562. package/docs/architecture/schema-inventory.md +8 -0
  563. package/docs/architecture-overview.md +70 -96
  564. package/docs/autonomy-and-human-roles.md +25 -0
  565. package/docs/cli/agent-usage.md +5 -3
  566. package/docs/cli/capability-routing.md +2 -2
  567. package/docs/cli/reference.md +13 -5
  568. package/docs/cockpit/README.md +4 -2
  569. package/docs/config.json +33 -23
  570. package/docs/conformance-acceptance.md +97 -0
  571. package/docs/conformance-aiwg-example.md +110 -0
  572. package/docs/context-engineering-contract.md +26 -0
  573. package/docs/context-management-patterns.md +2 -0
  574. package/docs/customization/README.md +8 -7
  575. package/docs/customization/project-local-quickstart.md +22 -0
  576. package/docs/development/architecture-illustrations.md +22 -0
  577. package/docs/development/cli-help-routing-audit.md +52 -0
  578. package/docs/development/devkit-overview.md +14 -13
  579. package/docs/development/skill-inventory.md +2 -2
  580. package/docs/docs-sources.json +27 -0
  581. package/docs/extensions/overview.md +11 -7
  582. package/docs/frameworks/forensics-complete/packet-evidence-integration.md +78 -0
  583. package/docs/frameworks/knowledge-base/overview.md +20 -2
  584. package/docs/frameworks/media-curator/overview.md +17 -6
  585. package/docs/frameworks/media-marketing-kit/overview.md +23 -15
  586. package/docs/frameworks/media-marketing-kit/quickstart.md +28 -8
  587. package/docs/frameworks/ops-complete/overview.md +17 -7
  588. package/docs/frameworks/ops-complete/packet-verification.md +53 -0
  589. package/docs/frameworks/ops-complete/quickstart.md +23 -3
  590. package/docs/frameworks/research-complete/overview.md +39 -26
  591. package/docs/frameworks/research-complete/packet-evidence.md +62 -0
  592. package/docs/frameworks/research-complete/quickstart.md +30 -12
  593. package/docs/frameworks/security-engineering/network-control-review.md +21 -0
  594. package/docs/getting-started/README.md +49 -116
  595. package/docs/getting-started/audit-existing-code.md +20 -21
  596. package/docs/getting-started/cognitive-walkthrough-1385-1386.md +2 -1
  597. package/docs/getting-started/daemon-and-automation.md +2 -1
  598. package/docs/getting-started/demo-script.md +2 -1
  599. package/docs/getting-started/existing-project.md +14 -38
  600. package/docs/getting-started/first-success-ask-steward.md +2 -1
  601. package/docs/getting-started/first-success-find-capability.md +2 -1
  602. package/docs/getting-started/first-success-start-intake.md +2 -1
  603. package/docs/getting-started/flow-and-gate-process.md +10 -5
  604. package/docs/getting-started/forensics-framework.md +13 -9
  605. package/docs/getting-started/install-connect-verify.md +16 -9
  606. package/docs/getting-started/just-try-it.md +40 -39
  607. package/docs/getting-started/key-addons.md +69 -19
  608. package/docs/getting-started/language-map.md +2 -1
  609. package/docs/getting-started/macos-install.md +2 -1
  610. package/docs/getting-started/marketing-framework.md +18 -11
  611. package/docs/getting-started/media-curator-framework.md +24 -12
  612. package/docs/getting-started/new-project.md +15 -28
  613. package/docs/getting-started/onboarding-research-refresh.md +2 -1
  614. package/docs/getting-started/onboarding-validation.md +2 -1
  615. package/docs/getting-started/prerequisites.md +17 -14
  616. package/docs/getting-started/provider-handoff.md +3 -3
  617. package/docs/getting-started/research-framework.md +24 -13
  618. package/docs/getting-started/scope-and-recovery.md +2 -1
  619. package/docs/getting-started/sdlc-framework.md +23 -14
  620. package/docs/getting-started/session-history.md +13 -0
  621. package/docs/getting-started/share-aiwg.md +2 -1
  622. package/docs/getting-started/start-here.md +47 -83
  623. package/docs/getting-started/storage-and-pkm.md +2 -1
  624. package/docs/getting-started/team-setup.md +34 -18
  625. package/docs/getting-started/verify-aiwg-is-working.md +12 -9
  626. package/docs/getting-started/writing-and-content.md +14 -10
  627. package/docs/how-it-works.md +86 -373
  628. package/docs/integrations/_manifest.json +1 -0
  629. package/docs/integrations/claude-code-quickstart.md +16 -13
  630. package/docs/integrations/codex-quickstart.md +17 -14
  631. package/docs/integrations/copilot-quickstart.md +17 -9
  632. package/docs/integrations/cross-platform-overview.md +43 -17
  633. package/docs/integrations/cursor-quickstart.md +17 -8
  634. package/docs/integrations/deepseek-harness-quickstart.md +40 -0
  635. package/docs/integrations/factory-quickstart.md +17 -8
  636. package/docs/integrations/hermes-quickstart.md +17 -8
  637. package/docs/integrations/openclaw-quickstart.md +17 -8
  638. package/docs/integrations/opencode-quickstart.md +17 -8
  639. package/docs/integrations/openhuman-quickstart.md +17 -8
  640. package/docs/integrations/pi-quickstart.md +15 -0
  641. package/docs/integrations/warp-terminal-quickstart.md +17 -8
  642. package/docs/integrations/windsurf-quickstart.md +18 -5
  643. package/docs/mcp/README.md +5 -2
  644. package/docs/models/hybrid-architectures.md +2 -0
  645. package/docs/network-analysis/compatibility.md +86 -0
  646. package/docs/overview/capabilities.md +61 -0
  647. package/docs/overview/executive-brief.md +48 -199
  648. package/docs/overview/reading-list.md +259 -0
  649. package/docs/overview/what-is-aiwg.md +71 -679
  650. package/docs/planning/session-intelligence/provider-conformance-matrix.json +11 -1
  651. package/docs/project-local/overview.md +9 -4
  652. package/docs/providers/deepseek-harness-sessions.md +30 -0
  653. package/docs/providers/deepseek-harness.md +197 -0
  654. package/docs/providers/indexed-access-audit.md +41 -0
  655. package/docs/providers/marketplace-consumer.md +2 -2
  656. package/docs/providers/provider-inventory.md +7 -3
  657. package/docs/quickstart-mmk.md +4 -3
  658. package/docs/quickstart-sdlc.md +3 -2
  659. package/docs/quickstart.md +28 -47
  660. package/docs/releases/v2026.9.5-announcement.md +97 -0
  661. package/docs/releases/v2026.9.6-announcement.md +71 -0
  662. package/docs/security/network-analysis-construction-gate.md +46 -0
  663. package/docs/security/network-analysis-threat-model.md +135 -0
  664. package/docs/sessions/cli.md +5 -1
  665. package/docs/sessions/exploration-validation.md +42 -0
  666. package/docs/skills/agent-skills.md +1 -0
  667. package/docs/storage/README.md +8 -6
  668. package/docs/storage/overview.md +24 -19
  669. package/docs/testing/uat/fortemi-live-dataset-uat-plan.md +32 -1
  670. package/docs/verification-contracts.md +15 -0
  671. package/docs/voice/channels.md +59 -0
  672. package/docs/voice/consumers.md +38 -0
  673. package/docs/voice/evaluation-protocol.md +78 -0
  674. package/docs/voice/evidence/channel-literals-development-2026-09-07.json +30 -0
  675. package/docs/voice/evidence/channel-literals-development-2026-09-07.md +25 -0
  676. package/docs/voice/evidence/channel-model-comparison-2026-09-07.json +88 -0
  677. package/docs/voice/evidence/channel-model-comparison-2026-09-07.md +29 -0
  678. package/docs/voice/evidence/core-guidance-development-2026-09-07.json +32 -0
  679. package/docs/voice/evidence/core-guidance-development-2026-09-07.md +23 -0
  680. package/docs/voice/evidence/current-model-expression-2026-09-07.json +99 -0
  681. package/docs/voice/evidence/current-model-expression-2026-09-07.md +26 -0
  682. package/docs/voice/evidence/expression-edits-development-2026-09-07.json +52 -0
  683. package/docs/voice/evidence/expression-edits-development-2026-09-07.md +29 -0
  684. package/docs/voice/evidence/frontier-feedback-ablation-2026-09-07.json +91 -0
  685. package/docs/voice/evidence/frontier-feedback-ablation-2026-09-07.md +34 -0
  686. package/docs/voice/evidence/gemma-native-schema-2026-09-07.json +19 -0
  687. package/docs/voice/evidence/gemma-native-schema-2026-09-07.md +23 -0
  688. package/docs/voice/evidence/harness-correction-2026-09-07.json +240 -0
  689. package/docs/voice/evidence/harness-correction-2026-09-07.md +28 -0
  690. package/docs/voice/evidence/harness-models-2026-09-07.json +37 -0
  691. package/docs/voice/evidence/harness-models-2026-09-07.md +30 -0
  692. package/docs/voice/evidence/model-eligibility.md +7 -0
  693. package/docs/voice/evidence/selector-development-2026-09-07.json +12764 -0
  694. package/docs/voice/evidence/selector-development-2026-09-07.md +32 -0
  695. package/docs/voice/evidence/span-repair-development-2026-09-07.json +27 -0
  696. package/docs/voice/evidence/span-repair-development-2026-09-07.md +11 -0
  697. package/docs/voice/evidence/tinystyler-development-2026-09-07.json +65 -0
  698. package/docs/voice/evidence/tinystyler-development-2026-09-07.md +15 -0
  699. package/docs/voice/exemplar-selection.md +91 -0
  700. package/docs/voice/fidelity.md +23 -0
  701. package/docs/voice/flows/README.md +11 -0
  702. package/docs/voice/flows/voice-critique-correction.flow.yaml +344 -0
  703. package/docs/voice/qualification.md +50 -0
  704. package/docs/voice/receipts-and-migration.md +19 -0
  705. package/docs/voice/revision.md +57 -0
  706. package/docs/voice/writer-profiles.md +63 -0
  707. package/docs/voice/writing-briefs.md +78 -0
  708. package/docs/welcome.html +49 -158
  709. package/docs/yaml-metalanguage.md +1 -1
  710. package/package.json +36 -6
  711. package/prebuilt/fortemi-core/framework/aiwg-fortemi-index-v2.json +1 -1
  712. package/prebuilt/fortemi-core/framework/manifest.json +2 -2
  713. package/schemas/catalog/catalog.json +3 -1
  714. package/schemas/catalog/domains/dataset.json +101 -0
  715. package/schemas/catalog/domains/network-analysis.json +121 -0
  716. package/schemas/catalog/domains/repository-json-schemas.json +20 -0
  717. package/schemas/catalog/domains/testing-quality.json +613 -0
  718. package/schemas/dataset/fortemi-live-qualification-receipt.v2.schema.json +196 -0
  719. package/schemas/dataset/fortemi-run-receipt/validation-1.0.1/authority.json +12 -0
  720. package/schemas/dataset/fortemi-run-receipt/validation-1.0.1/run-receipt.schema.json +819 -0
  721. package/schemas/mission-protocol/inventory-v1.json +33 -11
  722. package/schemas/network-analysis/analysis-recipe.v1.schema.json +279 -0
  723. package/schemas/network-analysis/governance-record.v1.schema.json +180 -0
  724. package/schemas/network-analysis/packet-evidence.v1.schema.json +487 -0
  725. package/test/fixtures/network-analysis/conformance-report.v1.json +62 -0
  726. package/test/fixtures/network-analysis/manifest.v1.json +68 -0
  727. package/tools/agents/deploy-agents.mjs +7 -3
  728. package/tools/agents/providers/antigravity.mjs +1 -1
  729. package/tools/agents/providers/base.mjs +3 -2
  730. package/tools/agents/providers/deepseek-harness.mjs +66 -0
  731. package/tools/agents/providers/hermes.mjs +1 -1
  732. package/tools/agents/providers/openhuman.mjs +2 -2
  733. package/tools/cli/doctor.mjs +1 -1
  734. package/tools/cli/validate-writing.mjs +20 -14
  735. package/tools/cli/wizard.mjs +1 -1
  736. package/tools/manifest/check-discovery-coverage.mjs +2 -2
  737. package/tools/providers/deepseek-harness-live-smoke.mjs +32 -0
  738. package/tools/providers/deepseek-harness-transport.mjs +358 -0
  739. package/tools/qualification/dataset-fortemi-execute.ts +70 -0
  740. package/tools/ralph-external/index.mjs +3 -3
  741. package/tools/ralph-external/lib/deepseek-harness-adapter.mjs +39 -0
  742. package/tools/ralph-external/lib/provider-adapter.mjs +3 -0
  743. package/tools/testing/audit-conformance-inventory.mjs +75 -0
  744. package/tools/testing/check-test-registration.mjs +73 -0
  745. package/tools/testing/conformance-example-state.mjs +17 -0
  746. package/tools/testing/run-conformance-example.mjs +128 -0
  747. package/tools/testing/verify-coverage-enforcement.mjs +61 -0
  748. package/tools/writing/writing-validator.mjs +34 -35
package/README.md CHANGED
@@ -1,12 +1,19 @@
1
1
  <div align="center">
2
2
 
3
- <a href="https://aiwg.io"><img src="https://aiwg.io/assets/badges/aiwg-hero-dark.png" alt="AIWG — multi-agent AI framework · one source of truth · 14 provider integrations including Google Antigravity CLI" width="680"></a>
3
+ <a href="https://aiwg.io"><img src="docs/.public/aiwg-readme-hero-v2.png" alt="AIWG — multi-agent AI framework, one
4
+ source of truth; network connecting AI tools" width="1000"></a>
4
5
 
5
6
  # AIWG
6
7
 
7
- **Multi-agent AI framework for 14 provider integrations, including Antigravity, Claude Code, Codex, Copilot, Cursor, Pi, and Warp**
8
+ **Reusable project context and specialist workflows for the AI tools you already use.**
8
9
 
9
- 200+ agents, 109+ CLI commands, 400+ deployable agent/skill/command/rule artifacts, 8 core frameworks, 32 addons, and a 40-plugin Claude Code marketplace. SDLC workflows, digital forensics, research management, marketing operations, media curation, ops infrastructure, knowledge base, and fine-tuning dataset curation — all deployable with one command.
10
+ Plan software, coordinate specialist reviews, prepare campaigns, investigate incidents, organize research, curate
11
+ media, and maintain operational knowledge. AIWG combines agents, skills, rules, templates, and workflow utilities
12
+ around these tasks, adapting them to your existing AI provider.
13
+
14
+ Project artifacts carry decisions from one session to the next. Domain frameworks supply the procedures; addons extend
15
+ them with writing profiles, task loops, memory, testing tools, and other capabilities. The sections below show what
16
+ you can do, how the pieces work together, and how to use them.
10
17
 
11
18
  The simplest setup is to paste this into a supported AI provider:
12
19
 
@@ -23,7 +30,7 @@ system with one self-verifying `aiwg use all` command. That command refreshes
23
30
  the indices, regenerates project context, verifies the resulting deployment,
24
31
  and reports whether a provider reload is actually required.
25
32
 
26
- For secure long-running agents, install AIWG Cockpit with a self-hosted Agentic
33
+ For long-running agents that need an isolated executor, optionally install AIWG Cockpit with a self-hosted Agentic
27
34
  Sandbox executor you control and audit:
28
35
 
29
36
  ```text
@@ -69,7 +76,7 @@ See [Web-Backed AIWG Resources](docs/install/web-backed-resources.md) for source
69
76
  selection, exact-version overrides, cache verification, offline use, and the
70
77
  current framework-graph constraints.
71
78
 
72
- Then ask your AI assistant to set up the project for AIWG. The agent-led setup
79
+ For a larger project, ask your AI assistant to establish project policy as well. The agent-led setup
73
80
  conversation should establish remotes, issue storage, delivery behavior,
74
81
  signing policy, and provider choices; the assistant may call `aiwg setup project`
75
82
  as the underlying CLI helper.
@@ -85,7 +92,7 @@ Agents and stewards setting up AIWG end-to-end should use the
85
92
  [![GitHub Stars](https://img.shields.io/github/stars/jmagly/aiwg?style=flat-square)](https://github.com/jmagly/aiwg/stargazers)
86
93
  [![Node Version](https://img.shields.io/badge/node-%E2%89%A520.0.0-brightgreen?style=flat-square&logo=node.js)](https://nodejs.org)
87
94
  [![TypeScript](https://img.shields.io/badge/TypeScript-5.x-blue?style=flat-square&logo=typescript)](https://www.typescriptlang.org)
88
- [![14 Providers](https://img.shields.io/badge/Providers-14-purple?style=flat-square)](#-platform-support)
95
+ [![15 Providers](https://img.shields.io/badge/Providers-15-purple?style=flat-square)](#platform-support)
89
96
  [![Listed on mcpservers.org](https://mcpservers.org/badge.svg)](https://mcpservers.org/servers/docs-aiwg-io)
90
97
 
91
98
  [![Built With AIWG](https://aiwg.io/assets/badges/built-with-aiwg-dark.png)](https://aiwg.io/badges)
@@ -125,15 +132,16 @@ same scoped rebuild command.
125
132
 
126
133
  If `npm install -g aiwg` fails with `EACCES` while writing to
127
134
  `/usr/local/lib/node_modules/aiwg`, npm is using a system-owned global install
128
- directory. The recommended Mac path is Node 24 through `nvm`, then:
135
+ directory. The [Node.js setup guide](docs/getting-started/install-node.md) covers the supported runtime and
136
+ version-manager choices. After setting up Node, run:
129
137
 
130
138
  ```bash
131
139
  npm install -g aiwg
132
140
  aiwg --version
133
141
  ```
134
142
 
135
- If Node is already installed and you need a quick recovery, use npm's current
136
- user-owned prefix guidance:
143
+ If Node is already installed and you need a quick recovery, one manual alternative is a user-owned npm prefix. Choose
144
+ this only after checking your existing Node version-manager configuration:
137
145
 
138
146
  ```bash
139
147
  npm config set prefix ~/.local
@@ -158,22 +166,33 @@ echo 'export PATH="$(npm config get prefix)/bin:$PATH"' >> ~/.zshrc # or ~/.ba
158
166
  source ~/.zshrc # or restart your shell
159
167
  ```
160
168
 
161
- You can also invoke AIWG without adjusting `PATH` by using `npx aiwg <command>`. For a broader health check — version, deployed providers, missing dependencies, kernel-skill probes — run `aiwg doctor`, which surfaces the same PATH guidance on every invocation if the binary isn't reachable.
169
+ You can also invoke AIWG without adjusting `PATH` by using `npx aiwg <command>`. For a broader health check — version,
170
+ deployed providers, missing dependencies, kernel-skill probes — run `aiwg doctor`. See the [troubleshooting
171
+ guide](docs/troubleshooting/index.md) for the full recovery paths.
162
172
 
163
173
  ---
164
174
 
165
175
  ## What AIWG Is
166
176
 
167
- AIWG is a deployment tool and support utility for AI context. At its core, `aiwg use` copies markdown and YAML source files into the paths each provider reads, so one source of truth works across 14 named provider integrations. A fifteenth `generic` adapter emits portable files for unrecognized or custom harnesses and is not counted as a named integration.
177
+ AIWG gives your AI assistant reusable project context and specialist workflows. Its deployment layer connects those
178
+ instructions to your provider: `aiwg use` copies markdown and YAML source files into the paths each provider reads, so
179
+ one source of truth works across 15 named provider integrations. A sixteenth `generic` adapter emits portable files
180
+ for unrecognized or custom harnesses and is not counted as a named integration.
168
181
 
169
- Around that core, AIWG ships agent-facing utilities for things the base platforms do not handle on their own: persistent artifact memory (`.aiwg/`), background orchestration, autonomous loops, artifact indexing, cost telemetry, health diagnostics, and more. These are tools the agent calls when you ask for something AIWG-shaped — you stay in chat. Most are opt-in. The deployment layer works standalone as plain text files the platform reads natively.
182
+ Around that core, AIWG ships agent-facing utilities for work that benefits from additional structure: persistent
183
+ artifact memory (`.aiwg/`), background orchestration, autonomous loops, artifact indexing, cost telemetry, health
184
+ diagnostics, and more. These are tools the agent calls when the task calls for them — you stay in chat. Most are
185
+ opt-in. The deployment layer works standalone as plain text files the platform reads natively.
170
186
 
171
187
  ### Project scope (recommended) vs user scope (global)
172
188
 
173
189
  `aiwg use` supports project deployments, additive user mirrors, and a
174
190
  user-global bootstrap:
175
191
 
176
- - **Project scope** — default. Run `aiwg use all --provider <provider>` from a project root and the artifacts land in that provider's project paths. One project's agent set never bleeds into another's session. **This is the recommended default for new users.**
192
+ - **Project scope** — default. Run `aiwg use all --provider <provider>` from a project root and the artifacts land in
193
+ that provider's project paths. This keeps project-specific instructions associated with the intended repository.
194
+ Some providers also use user-level surfaces; check the reported deployment scope. **This is the recommended default
195
+ for new users.**
177
196
  - **User scope (additive mirror)** — `aiwg use all --provider <provider> --scope user` keeps the
178
197
  project deployment and mirrors it to `~/.claude/agents/`,
179
198
  `~/.claude/skills/`, etc.
@@ -183,42 +202,59 @@ user-global bootstrap:
183
202
  Use `aiwg regenerate --provider <name>` to wire additional projects without
184
203
  deploying their own skill copies.
185
204
 
186
- The trade-off is real: when the same agent set loads into every session, context from one project can bleed into reasoning about another. Research (REF-720, *Lost in Multi-Turn Conversation*, MSR/Salesforce 2025) measured a 39% capability drop when this happens. The non-blocking project-isolation warning surfaces the trade-off at deploy time so the scope choice is informed. Neither scope is wrong; pick the one that fits the workflow.
205
+ Shared user-level instructions can be useful for personal conventions, while project-local instructions keep a team's
206
+ requirements and decisions with its repository. Review both scopes when a provider uses them together. The installer
207
+ and provider inventory describe where files will go, so you can distinguish instructions that follow you across
208
+ projects from instructions intended for this workspace.
187
209
 
188
210
  See the [Agentic Install Runbook](docs/agentic-install-runbook.md) for the
189
- zero-to-running setup path, and `https://github.com/jmagly/aiwg/blob/main/docs/cli/reference.md` (under `aiwg use` →
211
+ zero-to-running setup path, and the [CLI reference](docs/cli/reference.md) (under `aiwg use` →
190
212
  "Scope models") for the per-provider details and the global-install rough-edge
191
213
  inventory.
192
214
 
193
215
  ## Simple Building Blocks
194
216
 
195
- AIWG ships five primitive artifact types. All are plain text:
217
+ AIWG's workflow source is readable and editable. The main building blocks are:
196
218
 
197
- - **Agents** — specialized personas (Security Auditor, Test Architect) with a scoped toolset
198
- - **Skills** — natural-language workflows the platform auto-invokes on trigger phrases
199
- - **Commands** — explicit slash invocations (`/flow-security-review-cycle`)
200
- - **Rules** — enforcement directives the platform loads into every session
201
- - **Behaviors** — lifecycle hooks that fire on events (pre-write, post-session)
219
+ - **Agents** — specialist role instructions, such as Security Auditor or Test Architect, with defined responsibilities
220
+ and supported tool access.
221
+ - **Skills** — reusable procedures an agent can find from a goal and follow during a task.
222
+ - **Commands** — explicit ways to request a workflow through the provider or CLI.
223
+ - **Rules** — constraints for the assistant to follow, with tool-based checks where configured.
224
+ - **Behaviors** — lifecycle actions and hooks on providers that support them.
225
+ - **Templates** — structures for requirements, briefs, review reports, runbooks, and other outputs.
202
226
 
203
- Each is a single `.md` file with YAML frontmatter. Nothing executes until an AI platform reads it.
227
+ Many assets use Markdown with YAML metadata; hooks and utilities may also include executable scripts or structured
228
+ configuration. The provider determines how each asset is loaded or invoked. A role definition is not a separate model,
229
+ and a written rule is not proof that its constraint was enforced.
204
230
 
205
231
  ## Why It Compounds
206
232
 
207
- Because the primitives are text, they compose without runtime coordination:
233
+ The building blocks become more useful when workflows share their outputs:
234
+
235
+ - A requirements analyst records acceptance criteria that the test engineer can later use to review coverage.
236
+ - A security reviewer reads the same design decision as the implementation agent and records concerns against that decision.
237
+ - A campaign brief gives content writers a shared audience, message, and review criteria.
238
+ - A research note connects a source to the claim it supports, so a later synthesis can revisit the original evidence.
239
+ - A runbook carries verification and recovery steps from planning into an operational change.
208
240
 
209
- - One agent file becomes one member of a **180-agent SDLC team** that reviews architecture, tests, security, and compliance in parallel.
210
- - One skill becomes a **natural-language entry point** — "run security review" routes to the right multi-agent flow on every platform that supports skills.
211
- - One **framework** (SDLC, forensics, marketing) bundles dozens of agents + skills + rules + templates that cross-reference each other. Deploying a framework deploys a working multi-agent ecosystem.
212
- - The `.aiwg/` directory gives those agents a **shared memory** — artifacts from Monday's requirements session are read by Thursday's test design.
213
- - Flows orchestrate **Primary Author → Parallel Reviewers → Synthesizer → Archive** patterns that no single-prompt workflow can match.
241
+ Frameworks package these relationships: agents, skills, rules, and templates reference one another, while `.aiwg/`
242
+ holds the project-specific work they produce. A review can follow **Primary Author → Reviewers → Synthesizer →
243
+ Approval → Archive**, using parallel execution where the provider and task support it.
214
244
 
215
- The leverage is not in any one file. It is that hundreds of small files — each independently readable and editable — snap together into workflows that would otherwise take a bespoke agent platform to build.
245
+ For example, Monday's architecture review can become Thursday's implementation checklist. The second session needs to
246
+ read the saved artifact and check that it is still applicable, but the team has a concrete record to work from rather
247
+ than reconstructing the decision from conversation fragments.
216
248
 
217
- This is also where the research background lives. AIWG implements patterns from cognitive science (Miller 1956, Sweller 1988), multi-agent systems (Jacobs et al. 1991, MetaGPT, AutoGen), and software engineering (Cooper's stage-gate, FAIR Principles, W3C PROV) — applied as file conventions and deployment rules, not as a runtime you depend on.
249
+ Research on structured artifacts, multi-agent review, and recovery informs this design. The [research
250
+ foundations](#research-foundations) section preserves that background separately from claims about AIWG's own
251
+ performance.
218
252
 
219
253
  ## How You Actually Use AIWG
220
254
 
221
- The user surface is the conversation with your AI tool. You install AIWG, deploy a framework, and then talk to the agent normally — "help me start a project", "run a security review", "find me a deploy workflow." The agent does the AIWG-specific work for you.
255
+ The user surface is the conversation with your AI tool. You install AIWG, deploy a framework, and then talk to the
256
+ agent normally — "help me start a project", "run a security review", "find me a deploy workflow." The agent can
257
+ discover the appropriate procedure, perform the task, and report what it verified.
222
258
 
223
259
  The CLI exists mostly for the agent to call under the hood. The commands a user typically runs by hand are a short list:
224
260
 
@@ -230,53 +266,86 @@ The CLI exists mostly for the agent to call under the hood. The commands a user
230
266
  - `aiwg doctor` — health check
231
267
  - `aiwg refresh` — keep the install current
232
268
 
233
- Everything else is agent territory. Discovery (`aiwg discover`), artifact lookup (`aiwg show`), the index, agent loops, mission control, MCP — those are tools the agent invokes during a chat when you ask for something AIWG-shaped. You stay in the conversation; the agent handles the lookups, runs the loops, and reports back.
269
+ Advanced operators can also use the CLI directly; most everyday work can stay in chat. Discovery (`aiwg discover`),
270
+ artifact lookup (`aiwg show`), the index, agent loops, mission control, MCP — those are tools the agent invokes during
271
+ a chat when the task calls for them. The agent handles the lookups, executes the selected workflow, and reports its
272
+ result and limitations.
234
273
 
235
- Turn the agent-side tooling on (it's on by default once you `aiwg use`) when you want persistence, parallelism, or automation. Turn it off and the deployed agents, skills, and rules still work — they are still text files the platform reads natively.
274
+ Deployment connects the supported workflow surface. Optional servers, storage services, and automation paths have
275
+ their own prerequisites; installing assets does not mean all those services are running. You can start with one review
276
+ or document task, then enable additional utilities as the work requires.
236
277
 
237
278
  ## What AIWG Is Not
238
279
 
239
- - **Not a prompt library.** Prompts are the artifacts, not the product. The product is placing the right prompts where the platform finds them.
240
- - **Not an LLM runtime.** AIWG never calls a model. The AI platform you already use does that; AIWG configures what it sees.
241
- - **Not a framework you import into your app.** Nothing is imported at build time. Your project gets a `.aiwg/` directory (artifacts) and a few provider-specific context dirs (deployed copies). Delete them and your app is unchanged.
280
+ AIWG adds context, procedures, and support utilities around the AI tools you use. The deployment core writes
281
+ provider-readable assets; optional orchestration and integration components can invoke provider runtimes or other
282
+ configured services. The [architecture overview](docs/architecture-overview.md) distinguishes those components.
283
+
284
+ It does not replace your provider subscription, your application's runtime, or the review needed before using
285
+ generated work. Most workflows operate on project artifacts and source files rather than requiring your application to
286
+ import an AIWG library. Keep generated configuration and project artifacts under the same review discipline as other
287
+ repository changes.
242
288
 
243
289
  ## Who It's For
244
290
 
245
- If you have used AI coding assistants and thought "this is amazing for small tasks but falls apart on anything complex," AIWG is the missing infrastructure layer that scales AI assistance to multi-week projects.
291
+ AIWG is useful to individual developers, engineering teams, technical leaders, researchers, marketers, and operators
292
+ whose work spans several tasks or sessions. It helps when you need reusable instructions, a shared record of
293
+ decisions, or reviews from more than one perspective.
294
+
295
+ You can start at several levels: a small documentation review, a focused code audit, a campaign brief, an
296
+ investigation plan, or a complete development lifecycle. The [capability guide](docs/overview/capabilities.md) offers
297
+ task-based routes, while this README keeps the broader feature and workflow detail available below.
246
298
 
247
299
  ---
248
300
 
249
301
  ## What Problems Does AIWG Solve?
250
302
 
251
- Base AI assistants (Claude, GPT-4, Copilot without frameworks) have three fundamental limitations:
303
+ AI-assisted projects often need a deliberate way to carry context forward, recover from failed attempts, and make
304
+ review criteria explicit. AIWG provides workflows and artifacts for each of those needs.
252
305
 
253
- ### 1. No Memory Across Sessions
306
+ ### 1. Maintaining Context Across Sessions
254
307
 
255
- Each conversation starts fresh. The assistant has no idea what happened yesterday, what requirements you documented, or what decisions you made last week. You re-explain context every morning.
308
+ Useful decisions can become scattered between conversations, issues, and source files. AIWG workflows save project
309
+ outputs in `.aiwg/`, including requirements, architecture decisions, risk notes, test strategies, and campaign
310
+ material.
256
311
 
257
- **Without AIWG**: Projects stall as context rebuilding eats time. A three-month project requires continuity, not fresh starts every session.
312
+ A later task can consult the relevant artifact and link its own work back to it. The requirements analyst writes a use
313
+ case; the test engineer reads it to identify missing coverage; the implementation review checks whether the behavior
314
+ matches the acceptance criteria. A changed decision can be recorded and propagated through those relationships.
258
315
 
259
- **With AIWG**: The `.aiwg/` directory maintains 50-100+ interconnected artifacts across days, weeks, and months. Later phases build on earlier ones automatically because memory persists. Agents read prior work via `@-mentions` instead of regenerating from scratch.
316
+ The structure also helps an agent select context for a large project. Instead of treating every file as equally
317
+ relevant, a task can begin with a requirement, design record, or source note and follow its supporting references.
318
+ Artifact lookup and indexing utilities help locate those records as the collection grows.
260
319
 
261
- The segmented structure also makes large projects tractable. As code files grow, the project doesn't become harder to reason about — agents load only the slice of memory relevant to the current task (`@requirements/UC-001.md`, `@architecture/sad.md`, `@testing/test-plan.md`) rather than the entire codebase. Each subdirectory is a focused knowledge domain that fits comfortably in context, while cross-references keep everything connected.
320
+ The benefit depends on keeping artifacts current and actually consulting them. A saved file is a reusable source of
321
+ context, not a guarantee that every later response will use it correctly.
262
322
 
263
- The artifact index (`aiwg index`) takes this further. Without any tooling, agents often need to browse 3-6 documents before finding what they need. AIWG's structured artifacts reduce this to 2-3. With the index enabled, agents resolve artifact lookups in one query more often than not — a direct hit on the right requirement, architecture decision, or test case without browsing.
323
+ ### 2. Recovering from Failed Attempts
264
324
 
265
- ### 2. No Recovery Patterns
325
+ A failing test or incomplete task needs a diagnosis, not just another attempt with the same assumptions. Agent loops
326
+ support an execute-and-verify cycle that records failure information, adapts the next attempt, and stops at configured
327
+ limits or escalation conditions.
266
328
 
267
- When AI generates broken code or flawed designs, you manually intervene, explain the problem, and hope the next attempt works. There is no systematic learning from failures, no structured retry, no checkpoint-and-resume.
329
+ Use a loop for a bounded change with an observable completion criterion: fixing a regression, bringing a module under
330
+ test, or carrying out a migration plan. External loop tooling adds process and session recovery where supported. Its
331
+ usefulness depends on the provider, environment, task boundaries, and verification command; unattended execution is
332
+ not a guarantee of completion.
268
333
 
269
- **Without AIWG**: Research shows 47% of AI workflows produce inconsistent outputs without reproducibility constraints (R-LAM, Sureshkumar et al. 2026). Debugging is trial-and-error.
334
+ The record of attempts can help a later session understand what was tried and why it failed. That makes the recovery
335
+ process reviewable even when the agent needs human input or cannot complete the task.
270
336
 
271
- **With AIWG**: The agent loop implements closed-loop self-correction — execute, verify, learn from failure, adapt strategy, retry. External Ralph survives crashes and runs for 6-8+ hours autonomously. Debug memory accumulates failure patterns so the agent doesn't repeat mistakes.
337
+ ### 3. Making Quality Criteria Explicit
272
338
 
273
- ### 3. No Quality Gates
339
+ Different reviews ask different questions. A security review examines exposure and trust boundaries; a performance
340
+ review examines expected load and bottlenecks; a test review checks whether acceptance criteria are exercised; a
341
+ writing review checks audience, clarity, and support for claims.
274
342
 
275
- Base assistants optimize for "sounds plausible" not "actually works." A general assistant critiques security, performance, and maintainability simultaneously — poorly. No domain specialization, no multi-perspective review, no human approval checkpoints.
343
+ AIWG supplies specialist roles and workflows that separate these concerns, combine their findings, and record open
344
+ decisions. Phase gates can check whether the required artifacts and reviews are ready before the work advances.
345
+ Project policy determines who must approve a decision and which checks are required.
276
346
 
277
- **Without AIWG**: Production code ships without architectural review, security validation, or operational feasibility assessment.
278
-
279
- **With AIWG**: 162 specialized agents provide domain expertise — Security Auditor reviews security, Test Architect reviews testability, Performance Engineer reviews scalability. Multi-agent review panels with synthesis. Human-in-the-loop gates at every phase transition. Research shows 84% cost reduction keeping humans on high-stakes decisions versus fully autonomous systems (Agent Laboratory, Schmidgall et al. 2025).
347
+ Multiple reviewers can still share an error. The practical benefit is a clearer review procedure and a saved account
348
+ of the findings, with tests or other independent checks where available.
280
349
 
281
350
  ---
282
351
 
@@ -284,13 +353,17 @@ Base assistants optimize for "sounds plausible" not "actually works." A general
284
353
 
285
354
  ### 1. Memory — Structured Semantic Memory
286
355
 
287
- The `.aiwg/` directory is a persistent artifact repository storing requirements, architecture decisions, test strategies, risk registers, and deployment plans across sessions. This implements Retrieval-Augmented Generation patterns (Lewis et al., 2020) — agents retrieve from an evolving knowledge base rather than regenerating from scratch.
356
+ The `.aiwg/` directory is a persistent artifact repository storing requirements, architecture decisions, test
357
+ strategies, risk registers, and deployment plans across sessions. Artifacts provide retrievable project knowledge that
358
+ can ground later work in recorded decisions and sources.
288
359
 
289
- Each artifact is discoverable via `@-mentions` (e.g., `@.aiwg/requirements/UC-001-login.md`). Context sharing between agents happens through artifacts: the requirements analyst writes use cases, the architecture designer reads them.
360
+ Artifacts can be referenced with `@-mentions` (e.g., `@.aiwg/requirements/UC-001-login.md`). Context sharing between
361
+ agents happens through artifacts: the requirements analyst writes use cases, the architecture designer reads them.
290
362
 
291
363
  ### 2. Reasoning — Multi-Agent Deliberation with Synthesis
292
364
 
293
- Instead of a single general-purpose assistant, AIWG provides 162 specialized agents organized by domain. Complex artifacts go through multi-agent review panels:
365
+ AIWG provides specialist role definitions organized by domain. A workflow can route a complex artifact through the
366
+ reviewers the task requires:
294
367
 
295
368
  ```
296
369
  Architecture Document Creation:
@@ -304,11 +377,14 @@ Architecture Document Creation:
304
377
  4. Human approval gate → accept, iterate, or escalate
305
378
  ```
306
379
 
307
- Research shows 17.9% accuracy improvement with multi-path review on complex tasks (Wang et al., GSM8K benchmarks, 2023). Agent specialization means security review is done by a security specialist, not a generalist.
380
+ Each reviewer receives a defined responsibility and relevant context. The synthesis should resolve duplicate or
381
+ conflicting findings, preserve uncertainty, and identify which conclusions were checked against sources or tests.
382
+ Parallel reviews require the corresponding provider capability and available task budget.
308
383
 
309
384
  ### 3. Learning — Closed-Loop Self-Correction (Ralph)
310
385
 
311
- Ralph executes tasks iteratively, learns from failures, and adapts strategy based on error patterns. Research from Roig (2025) shows recovery capability — not initial correctness — predicts agentic task success.
386
+ Ralph executes tasks iteratively and uses verification results to guide the next attempt. Its task record can preserve
387
+ failure analysis and revised strategies for subsequent iterations.
312
388
 
313
389
  ```
314
390
  Ralph Iteration:
@@ -316,14 +392,16 @@ Ralph Iteration:
316
392
  2. Verify results (tests pass, lint clean, types check)
317
393
  3. If failure: analyze root cause → extract structured learning → adapt strategy
318
394
  4. Log iteration state (checkpoint for resume)
319
- 5. Repeat until success or escalate to human after 3 failed attempts
395
+ 5. Repeat within configured limits; stop or escalate when required
320
396
  ```
321
397
 
322
- External Ralph adds crash resilience: PID file tracking, automatic restart, cross-session persistence. Tasks run for 6-8+ hours surviving terminal disconnects and system reboots.
398
+ External Ralph adds process tracking, session persistence, and recovery controls. For long-running work, define a time
399
+ or iteration budget and inspect the provider-specific recovery behavior; surviving a particular failure depends on how
400
+ the runner and host are configured.
323
401
 
324
402
  ### 4. Verification — Bidirectional Traceability
325
403
 
326
- AIWG maintains links between documentation and code to ensure artifacts stay synchronized:
404
+ AIWG supports links between documentation and code so reviewers can inspect relationships and find drift:
327
405
 
328
406
  ```typescript
329
407
  // src/auth/login.ts
@@ -335,7 +413,9 @@ AIWG maintains links between documentation and code to ensure artifacts stay syn
335
413
  export function authenticateUser(credentials: Credentials): Promise<AuthResult> {
336
414
  ```
337
415
 
338
- Verification types: Doc → Code, Code → Doc, Code → Tests, Citations → Sources. The retrieval-first citation architecture reduces citation hallucination from 56% to 0% (LitLLM benchmarks, ServiceNow 2025).
416
+ Verification can follow Doc → Code, Code → Doc, Code → Tests, and Citations → Sources. These relationships make claims
417
+ easier to inspect; they do not prove that implementation or citations are correct. Ask the workflow to report missing
418
+ targets, inconsistent behavior, and unsupported source claims.
339
419
 
340
420
  ### 5. Planning — Phase Gates with Cognitive Load Management
341
421
 
@@ -346,16 +426,16 @@ Inception → Elaboration → Construction → Transition → Production
346
426
  LOM ABM IOC PR
347
427
  ```
348
428
 
349
- Cognitive load optimization follows Miller's 7±2 limits (1956) and Sweller's worked examples approach (1988):
350
-
351
- - 4 phases (not 12)
352
- - 3-5 artifacts per phase (not 20)
353
- - 5-7 section headings per template (not 15)
354
- - 3-5 reviewers per panel (not 10)
429
+ Phase structure gives a team bounded decisions and review points: establish goals during Inception, evaluate design
430
+ and risks during Elaboration, implement and verify during Construction, and prepare operational handoff during
431
+ Transition. Templates help make the expected outputs visible. The project determines which artifacts, reviewers, and
432
+ approval gates are appropriate; a small task need not use the full lifecycle.
355
433
 
356
434
  ### 6. Style — Controllable Voice Generation
357
435
 
358
- Voice profiles provide continuous control over AI writing style using 12 parameters (formality, technical depth, sentence variety, jargon density, personal tone, humor, directness, examples ratio, uncertainty acknowledgment, opinion strength, transition style, authenticity markers).
436
+ Voice profiles describe writing preferences such as formality, technical depth, directness, examples, uncertainty, and
437
+ sentence variation. They give the assistant a reusable style specification that can be reviewed against the audience
438
+ and document purpose.
359
439
 
360
440
  Built-in voices: `technical-authority` (docs, RFCs), `friendly-explainer` (tutorials), `executive-brief` (summaries), `casual-conversational` (blogs, social). Create custom voices from your existing content with `/voice-create`.
361
441
 
@@ -363,7 +443,13 @@ Built-in voices: `technical-authority` (docs, RFCs), `friendly-explainer` (tutor
363
443
 
364
444
  ## A Real Project Walkthrough
365
445
 
366
- Here is how the six components work together across a project lifecycle. How long each phase takes depends entirely on the project — AIWG is a force multiplier, not a clock. Most projects arrive at a complete, reviewed document set in hours to a day. What takes time is the human work that matters: reviewing, editing, and making decisions. The more input your team provides, the better the output. AIWG memory lets operators participate through the tools they already use — industry-standard documents and templates, issues, and knowledge bases.
446
+ The following illustrative customer-portal project shows how the components connect across a lifecycle. The commands
447
+ are provider-facing workflow examples, not a report of a measured project outcome. Natural-language requests can
448
+ select the same procedures when the provider does not expose the shown slash syntax.
449
+
450
+ Start with a bounded goal, agree on acceptance criteria, and inspect the artifacts at each step. The time and review
451
+ effort depend on the project, source quality, tools, and decisions involved. Smaller changes can enter at the phase
452
+ that fits their current state rather than repeating the entire lifecycle.
367
453
 
368
454
  ### Inception
369
455
 
@@ -409,25 +495,31 @@ Here is how the six components work together across a project lifecycle. How lon
409
495
  ```
410
496
 
411
497
  **Planning**: Deployment checklist — monitoring, rollback plan, incident response
412
- **Learning**: Ralph retries deployment steps if validation fails
498
+ **Learning**: Failed validation produces a diagnosis and a revised plan; retries follow the operation's recovery and
499
+ approval requirements
413
500
  **Verification**: Deployment scripts reference architecture (which services, what order)
414
501
  **Human Gate**: Operations team reviews deployment plan → approves production release
415
502
 
416
503
  ---
417
504
 
418
- ## Quantified Claims and Evidence
505
+ ## Claims, Evaluation, and Evidence
419
506
 
420
- AIWG makes specific, falsifiable claims backed by peer-reviewed research:
507
+ Evaluate a workflow against the task it is meant to support. AIWG provides structures for review, traceability, and
508
+ recovery; it does not promise a fixed cost saving, perfect citations, or error-free execution.
421
509
 
422
- | Claim | Evidence | Source |
423
- |-------|----------|--------|
424
- | 84% cost reduction with human-in-the-loop vs fully autonomous | Agent Laboratory study | Schmidgall et al. (2025) |
425
- | 47% workflow failure rate without reproducibility constraints | R-LAM evaluation | Sureshkumar et al. (2026) |
426
- | 0% citation hallucination with retrieval-first vs 56% generation-only | LitLLM benchmarks | ServiceNow (2025) |
427
- | 17.9% accuracy improvement with multi-path review | GSM8K benchmarks | Wang et al. (2023) |
428
- | 18.5x improvement with tree search on planning tasks | Game of 24 results | Yao et al. (2023) |
510
+ | Capability to evaluate | Evidence to collect | Useful comparison |
511
+ |------------------------|---------------------|-------------------|
512
+ | Persistent project context | Whether a later session reads and applies a prior decision | The same kind of task with the team's usual handoff |
513
+ | Specialist review | Correct findings, missed issues, and reviewer effort | A comparable review using the existing process |
514
+ | Task recovery | Failure diagnosis, attempts, limits, and final verification | Similar failures handled without the loop |
515
+ | Citation checking | Source existence and support for each material claim | Manual source inspection |
516
+ | Structured planning | Completeness and usefulness of the resulting plan | The team's existing planning artifact |
517
+ | Cost and throughput | Model calls, elapsed time, and human review effort | Comparable tasks with the same acceptance criteria |
429
518
 
430
- Full references: [docs/research/](docs/research/)
519
+ Research results from Agent Laboratory, self-consistency, tree search, and other systems inform the design. Their
520
+ percentages and benchmark scores are not measurements of AIWG. The [research foundations](#research-foundations) and
521
+ [reading list](docs/overview/reading-list.md) retain the underlying sources. The [executive
522
+ brief](docs/overview/executive-brief.md) describes a practical pilot.
431
523
 
432
524
  ---
433
525
 
@@ -441,13 +533,17 @@ Multi-week or multi-month projects where requirements evolve, multiple stakehold
441
533
 
442
534
  ### Not the Best Fit
443
535
 
444
- Single-session tasks where no memory is needed, quality gates are overkill, and overhead exceeds value.
536
+ A full lifecycle is usually unnecessary for a one-off question that needs no shared context, saved artifact, or
537
+ follow-up. A focused writing, lookup, or review capability can still be useful without introducing every phase and
538
+ gate.
445
539
 
446
540
  **Examples**: "Write a Python script to parse this CSV," "Fix this typo," "Explain how this code works."
447
541
 
448
542
  ### The Trade-off
449
543
 
450
- AIWG adds structure (templates, phases, gates) that slows trivial tasks but scales to complex multi-week workflows. If your project fits in a single conversation, use a base assistant. If it spans days, weeks, or months, AIWG provides the infrastructure to maintain quality and context.
544
+ Match the workflow to the task. Use a bounded skill for a small result, a saved artifact when work needs to carry
545
+ forward, and a phase-based process when coordination and review warrant it. Additional context, reviewers, or
546
+ verification steps can add model calls and human effort; judge their value against the outcome you need.
451
547
 
452
548
  ```
453
549
  User intent → AIWG CLI → Deploy agents + rules + templates → AI platform
@@ -457,11 +553,11 @@ User intent → AIWG CLI → Deploy agents + rules + templates → AI platform
457
553
  │ Cursor / Warp / Factory /
458
554
  ▼ OpenCode / Codex / Devin Desktop
459
555
  ┌──────────────┐
460
- │ 188 Agents │ Specialized AI personas with domain expertise
461
- │ 50 Commands │ CLI + slash commands for workflow automation
462
- │ 128 Skills │ Natural language workflow triggers
463
- │ 35 Rules │ Enforcement patterns (security, quality, anti-laziness)
464
- │ 334 Templates│ SDLC artifact templates with progressive disclosure
556
+ │ Agents │ Specialized AI personas with domain expertise
557
+ │ Commands │ CLI + slash commands for workflow automation
558
+ │ Skills │ Natural language workflow triggers
559
+ │ Rules │ Enforcement patterns (security, quality, anti-laziness)
560
+ │ Templates │ SDLC artifact templates with progressive disclosure
465
561
  └──────────────┘
466
562
  │
467
563
  ▼
@@ -474,17 +570,19 @@ User intent → AIWG CLI → Deploy agents + rules + templates → AI platform
474
570
 
475
571
  > For visual diagrams of AIWG's architecture, deploy flow, and discovery model, see [`docs/architecture-overview.md`](docs/architecture-overview.md). The prose walkthrough lives in [`docs/how-it-works.md`](docs/how-it-works.md).
476
572
 
477
- **At a glance** — AIWG is a deploy-time tool. `aiwg use` copies plain-text files into your AI platform's native directories and exits. Nothing runs in the background; the AI platform's own loader handles everything from there.
573
+ **At a glance** — the deployment layer copies instructions into provider-readable locations, connects project context,
574
+ and reports verification. The provider loads those instructions. Optional runtime components, such as artifact
575
+ services or orchestration tools, perform additional work when configured and invoked.
478
576
 
479
577
  ```mermaid
480
578
  flowchart LR
481
579
  subgraph Source["AIWG framework source"]
482
580
  direction TB
483
- KERN[25 kernel skills<br/>within provider listing budgets]
484
- STD[~455 standard skills<br/>read from $AIWG_ROOT]
485
- AGENT[200+ agents]
486
- RULES[60+ rules]
487
- TPL[100+ templates]
581
+ KERN[Kernel skills<br/>within provider listing budgets]
582
+ STD[Standard skills<br/>read from $AIWG_ROOT]
583
+ AGENT[Specialist agents]
584
+ RULES[Workflow rules]
585
+ TPL[Artifact templates]
488
586
  end
489
587
 
490
588
  CLI([aiwg use all<br/>--provider X]) --> DEPLOY
@@ -538,22 +636,41 @@ The orchestration pattern: **Primary Author → Parallel Reviewers → Synthesiz
538
636
 
539
637
  ## Features
540
638
 
541
- - **188 specialized agents** — domain experts across testing, security, architecture, DevOps, cloud, frontend, backend, data engineering, documentation, and more
542
- - **50 CLI commands** — framework deployment, project scaffolding, iterative execution, metrics, reproducibility validation
543
- - **128 workflow skills** — natural language triggers for regression testing, forensics, voice profiles, quality gates, and CI/CD integration
544
- - **35 enforcement rules** — anti-laziness detection, token security, citation integrity, executable feedback, failure mitigation across 6 LLM archetypes
545
- - **334 artifact templates** — progressive disclosure templates for requirements, architecture, testing, security, deployment, and more
546
- - **Multi-provider support** — deploy to Google Antigravity CLI, Claude Code, OpenAI Codex, GitHub Copilot, Cursor, Factory AI, Hermes, OpenCode, OpenClaw, OpenHuman, [Pi Coding Agent](https://pi.dev/), Oh My Pi, Warp Terminal, and Devin Desktop
547
- - **8 core frameworks + training marketplace package** — SDLC, Digital Forensics, Marketing Operations, Research Management, Media Curation, Ops Infrastructure, Knowledge Base, Security Engineering, plus [`aiwg-training`](https://github.com/jmagly/aiwg-training) for fine-tuning dataset curation (corpus-to-dataset pipeline with DPO/KTO/ORPO/SimPO export)
548
- - **32 addons** — compound memory, line memory, llm-wiki (Obsidian-native knowledge base), RLM recursive decomposition, fleet operations, browser control, testing quality, and more
549
- - **40 Claude Code plugins** — the complete framework and addon catalog is installable independently from the AIWG marketplace
550
- - **Agent Loop** — iterative task execution with automatic error recovery and crash resilience (6-8 hour sessions)
551
- - **RLM addon** — recursive context decomposition for processing 10M+ tokens via sub-agent delegation
552
- - **YAML metalanguage** — declarative schema-validated workflow definitions (JSON Schema 2020-12)
553
- - **MCP server** — Model Context Protocol integration for tool-based AI workflows
554
- - **Bidirectional traceability** — @-mention system linking requirements → architecture → code → tests
555
- - **FAIR-aligned artifacts** — W3C PROV provenance, GRADE quality assessment, persistent REF-XXX identifiers
556
- - **Reproducibility validation** — deterministic execution modes, checkpoints, configuration snapshots
639
+ - **Specialist agents** — roles for architecture, implementation, testing, security, cloud, data engineering,
640
+ research, content, and operations.
641
+ - **Workflow skills and commands** — discoverable procedures for reviews, intake, research, curation, planning, and delivery.
642
+ - **Rules and review criteria** — instructions for preserving work, handling sensitive configuration, checking claims,
643
+ and reporting verification.
644
+ - **Artifact templates** — structured requirements, design decisions, campaign briefs, source notes, runbooks, and
645
+ review reports.
646
+ - **Multi-provider deployment** — Google Antigravity CLI, Claude Code, OpenAI Codex, GitHub Copilot, Cursor, DeepSeek
647
+ Harness, Factory AI, Hermes, OpenCode, OpenClaw, OpenHuman, Pi Coding Agent, Oh My Pi, Warp Terminal, and Devin
648
+ Desktop.
649
+ - **Domain frameworks** — software development, forensics, marketing, research, media curation, operations, knowledge
650
+ base, and security engineering.
651
+ - **Dataset workflows** — assessment, indexing, lineage, synchronization, and retirement through dataset intelligence;
652
+ the separate [`aiwg-training`](https://github.com/jmagly/aiwg-training) project covers training-data curation and
653
+ exports.
654
+ - **Memory addons** — compound memory, line memory, wiki-oriented knowledge, and artifact lookup for different
655
+ persistence needs.
656
+ - **Writing and voice tools** — reusable voice profiles, context-sensitive diagnostics, revision workflows, and
657
+ alternatives for content generation.
658
+ - **Testing quality** — test conformance, reversible normalization, mutation testing, and flaky-test review.
659
+ See the [AIWG test conformance example](docs/conformance-aiwg-example.md) for reviewed source controls
660
+ and runner evidence.
661
+ - **Agent loops** — bounded execution, failure analysis, checkpoints, and supported process recovery.
662
+ - **RLM** — recursive context decomposition for tasks whose source material needs to be divided into smaller working sets.
663
+ - **YAML metalanguage** — structured workflow and artifact definitions with schema-oriented validation.
664
+ - **MCP integration** — tools and resources exposed through configured servers and provider connections.
665
+ - **Traceability and provenance** — relationships between requirements, code, tests, sources, and generated artifacts.
666
+ - **Session history and diagnostics** — import and inspect prior AI work, check deployment health, and diagnose
667
+ provider wiring.
668
+ - **Project-local extensions and marketplace delivery** — keep custom instructions with the project and package
669
+ reusable capabilities through the appropriate distribution path.
670
+
671
+ The [framework and addon catalog](#what-you-get) below describes these capabilities in more detail. Compatibility and
672
+ execution requirements are explicit in the [provider inventory](docs/providers/provider-inventory.md) and [CLI
673
+ reference](docs/cli/reference.md).
557
674
 
558
675
  ---
559
676
 
@@ -561,53 +678,69 @@ The orchestration pattern: **Primary Author → Parallel Reviewers → Synthesiz
561
678
 
562
679
  > **Prerequisites:** Node.js >=20.0.0 and an AI platform (Claude Code, GitHub Copilot, Cursor, Warp Terminal, or others). New installs should prefer Node 24. See [Prerequisites Guide](docs/getting-started/prerequisites.md) for details.
563
680
 
564
- > **Verifying releases (v2026.5.3+):** Every AIWG release ships with Sigstore-anchored npm provenance, a signed git tag, a cosign keyless tarball signature, and a signed CycloneDX SBOM. Verification is optional but recommended:
565
- >
566
- > ```bash
567
- > npm view aiwg@2026.5.3 --json | jq .dist.attestations
568
- > ```
569
- >
570
- > Full walkthrough at [`docs/releases/verifying.md`](docs/releases/verifying.md). Adopt the same pattern for your own packages: [`docs/security/supply-chain-hardening.md`](docs/security/supply-chain-hardening.md).
681
+ > **Release verification:** Inspect the provenance and signature material for the release you install. The
682
+ [verification guide](docs/releases/verifying.md) describes the available artifacts and commands.
571
683
 
572
684
  ### Install & Deploy
573
685
 
686
+ The prompt-led installer at the top of this README is the canonical beginner path. For manual setup:
687
+
574
688
  ```bash
575
- # Install globally
576
- npm install -g aiwg
689
+ npm i -g aiwg
690
+ cd /path/to/your/project
691
+ aiwg use all --provider claude # replace claude with your provider selector
692
+ ```
577
693
 
578
- # Deploy to your project
579
- cd your-project
580
- aiwg use sdlc # Full SDLC framework (98 agents, 38 rules, 200+ templates)
581
- aiwg use forensics # Digital forensics & incident response (13 agents, 10 skills)
582
- aiwg use marketing # Marketing operations (37 agents, 87+ templates)
583
- aiwg use media-curator # Media archive management (6 agents, 9 commands)
584
- aiwg use research # Research workflow automation (8 agents, 8-stage pipeline)
585
- aiwg use civic-action # Cited civic review and publication-preparation kit
586
- aiwg use rlm # RLM addon (recursive context decomposition)
587
- aiwg use all # Everything
588
-
589
- # Existing projects: preview, then transactionally extract the canonical graph
590
- aiwg regenerate --existing-project --dry-run
591
- aiwg regenerate --existing-project --apply
592
- aiwg workspace-context doctor
694
+ Deployment refreshes the shared context and reports verification and any required reload. Follow that result, then ask
695
+ the agent to check the intended project and its AIWG connection. The [manual installation
696
+ reference](docs/cli/install-and-repair.md) covers the terminal path in detail.
593
697
 
594
- # Recommended default: inspect project state and select the safe branch
698
+ For a deliberately narrower deployment, choose the relevant framework or addon instead of `all`. These are
699
+ alternatives, not a sequence of required setup steps:
700
+
701
+ ```bash
702
+ aiwg use sdlc --provider claude # Software development
703
+ aiwg use forensics --provider claude # Investigation workflows
704
+ aiwg use marketing --provider claude # Campaign and content work
705
+ aiwg use media-curator --provider claude # Media collections
706
+ aiwg use research --provider claude # Research artifacts
707
+ aiwg use civic-action --provider claude # Civic review and preparation
708
+ aiwg use rlm --provider claude # Context decomposition
709
+ ```
710
+
711
+ For maintenance or an existing workspace that needs context migration, preview the relevant regeneration branch rather
712
+ than treating every branch as installation:
713
+
714
+ ```bash
595
715
  aiwg regenerate --dry-run
596
716
  aiwg regenerate
597
717
 
598
- # Fresh or already-migrated projects: ordinary canonical refresh
599
- aiwg regenerate --workspace
718
+ # Existing-project extraction, when that is the intended operation:
719
+ aiwg regenerate --existing-project --dry-run
720
+ aiwg regenerate --existing-project --apply
721
+ aiwg workspace-context doctor
722
+ ```
723
+
724
+ The [regeneration guide](docs/regenerate-guide.md) also covers canonical refresh and legacy compatibility. Use the
725
+ branch that matches the workspace state. To scaffold a new project rather than connect the current one, see the
726
+ [new-project guide](docs/getting-started/new-project.md).
600
727
 
601
- # Legacy compatibility only: inline AIWG context in provider startup files
602
- aiwg regenerate --full-inject
728
+ ### Get a First Useful Result
603
729
 
604
- # Or scaffold a new project
605
- aiwg new my-project
730
+ After setup, ask your agent:
606
731
 
607
- # Check installation health
608
- aiwg doctor
732
+ ```text
733
+ Use AIWG to review this project's README for unclear positioning and missing
734
+ onboarding steps. Save a report at
735
+ .aiwg/marketing/brand/audit/readme-review.md with file references and the
736
+ three highest-priority fixes. Leave the README unchanged.
609
737
  ```
610
738
 
739
+ Open the report and check the source references, reader impact, and proposed fixes. In a later session, ask the agent
740
+ to read that report and implement the first agreed change. The [first-result
741
+ walkthrough](docs/getting-started/just-try-it.md) includes an illustrative finding and alternative tasks. Setup
742
+ readiness is the prerequisite; a useful artifact is what lets you assess the workflow.
743
+
611
744
  ### Customize Without Forking
612
745
 
613
746
  Author project-specific rules, skills, agents, addons, or frameworks
@@ -637,7 +770,7 @@ The bundle is **byte-identical** in shape to its upstream form, so
637
770
  /plugin install compound-memory@aiwg
638
771
  ```
639
772
 
640
- The marketplace contains 40 independently packaged framework and addon
773
+ The marketplace contains independently packaged framework and addon
641
774
  plugins, so you can install only the capabilities a Claude Code workspace
642
775
  needs. Source-distributed opt-in addons such as Civic Action deploy with
643
776
  `aiwg use civic-action` and do not imply a marketplace wrapper.
@@ -657,6 +790,8 @@ aiwg use all --provider devin # Devin Desktop
657
790
  aiwg use all --provider openclaw # OpenClaw
658
791
  aiwg use all --provider hermes # Hermes
659
792
  aiwg use all --provider openhuman # OpenHuman
793
+ aiwg use all --provider pi # Pi Coding Agent
794
+ aiwg use all --provider omp # Oh My Pi
660
795
  ```
661
796
 
662
797
  `all` means the complete deployable end-user surface. It intentionally omits
@@ -665,11 +800,13 @@ directly.
665
800
 
666
801
  ### First-Party Integrators
667
802
 
668
- Some partners ship AIWG bundled in their runtime — no `aiwg use` step required. Install the partner tool and AIWG is already wired up.
803
+ AIWG can also be distributed through another runtime. Check the integrator's version and included assets before
804
+ assuming that its bundled surface matches a standalone AIWG install. Follow that runtime's setup instructions, then
805
+ verify the project connection.
669
806
 
670
807
  | Partner | Install | What you get |
671
808
  |---------|---------|--------------|
672
- | **[Omnius](https://www.npmjs.com/package/omnius)** | `npm i -g omnius` | AIWG framework set, skills, agents, and rules embedded in the Omnius autonomous coding agent. Discoverable through `aiwg discover` from inside Omnius sessions, surfaceable through Omnius's MCP and REST bridges. |
809
+ | **[Omnius](https://www.npmjs.com/package/omnius)** | `npm i -g omnius` | An integration path for AIWG assets in an autonomous coding runtime. Consult the package documentation for the bundled version, supported asset surface, and setup requirements. |
673
810
 
674
811
  If you ship a product that bundles AIWG and want to be listed here, open an issue at https://github.com/jmagly/aiwg/issues.
675
812
 
@@ -677,163 +814,200 @@ If you ship a product that bundles AIWG and want to be listed here, open an issu
677
814
 
678
815
  ## What You Get
679
816
 
680
- ### Frameworks (8)
681
-
682
- | Framework | Agents | Templates | What It Does |
683
- |-----------|--------|-----------|--------------|
684
- | **[SDLC Complete](agentic/code/frameworks/sdlc-complete/)** | 98 | 200+ | Full software development lifecycle — Inception through Production with multi-agent orchestration, quality gates, and DORA metrics |
685
- | **[Forensics Complete](agentic/code/frameworks/forensics-complete/)** | 13 | 8 | Digital forensics and incident response — evidence acquisition, timeline reconstruction, IOC extraction, Sigma rule hunting. NIST SP 800-86, MITRE ATT&CK, STIX 2.1 |
686
- | **[Media/Marketing Kit](agentic/code/frameworks/media-marketing-kit/)** | 37 | 87+ | End-to-end marketing operations — strategy, content creation, campaign management, brand compliance, analytics, and reporting |
687
- | **[Media Curator](agentic/code/frameworks/media-curator/)** | 6 | — | Intelligent media archive management — discography analysis, source discovery, quality filtering, transcript sidecars, research handoff preparation, multi-platform export (Plex, Jellyfin, MPD) |
688
- | **[Research Complete](agentic/code/frameworks/research-complete/)** | 8 | 6 | Academic research automation — paper discovery, citation management, RAG-based summarization, GRADE quality scoring, FAIR compliance, W3C PROV provenance |
689
- | **[Knowledge Base](agentic/code/frameworks/knowledge-base/)** | — | 5 | General-purpose LLM-assisted wiki — source ingest, entity/concept pages, source summaries, comparisons, syntheses, health checks, and emergent taxonomy |
690
- | **[Ops Complete](agentic/code/frameworks/ops-complete/)** | 12 | 3 | Operational infrastructure — incident management, runbooks, troubleshooting workflows |
691
- | **[Security Engineering](agentic/code/frameworks/security-engineering/)** | 2 | 5 | Applied security beyond STRIDE/OWASP — cryptographic primitive selection, chain-of-trust integrity, authentication-factor architecture, degraded-mode design, runtime secret hygiene, supply-chain trust, physical-access threats. Pattern-based, product-agnostic |
692
-
693
- ### Addons (32)
694
-
695
- | Addon | What It Does |
696
- |-------|--------------|
697
- | **[Civic Action](agentic/code/addons/civic-action/)** | Cited, opt-in decision support for source review, records planning, meeting reconciliation, public-technology research, local-resource profiles, corrections, and publication preparation |
698
- | **[RLM](agentic/code/addons/rlm/)** | Recursive context decomposition — process 10M+ tokens via sub-agent delegation with parallel fan-out |
699
- | **[Writing Quality](agentic/code/addons/writing-quality/)** | Content validation, AI pattern detection, authentic voice enforcement |
700
- | **[Testing Quality](agentic/code/addons/testing-quality/)** | TDD enforcement, mutation testing, flaky test detection and repair |
701
- | **[Voice Framework](agentic/code/addons/voice-framework/)** | 4 built-in voice profiles (technical-authority, friendly-explainer, executive-brief, casual-conversational) with create/blend/apply skills |
702
- | **[UAT-MCP Toolkit](agentic/code/addons/uat-mcp/)** | User acceptance testing with MCP-powered test execution, coverage tracking, and regression detection |
703
- | **[AIWG Evals](agentic/code/addons/aiwg-evals/)** | Agent evaluation framework — archetype resistance testing (Roig 2025), performance benchmarks, quality scoring |
704
- | **[Agent Loop](agentic/code/addons/agent-loop/)** | Iterative task execution engine (`aiwg ralph` / `aiwg agent-loop-ext`) — automatic error recovery, crash resilience, completion tracking |
705
- | **[Agentic Installer](agentic/code/addons/agentic-installer/)** | `setup.aiwg.io/v1` SetupManifest installer — cross-platform install workflows with recovery |
706
- | **[AIWG Dev](agentic/code/addons/aiwg-dev/)** | AIWG development tooling — extension scaffolding, local-source dev mode |
707
- | **[Daemon](agentic/code/addons/daemon/)** | Persistent daemon mode — background sessions, task queue, health monitoring |
708
- | **[Compound Memory](agentic/code/addons/compound-memory/)** | Persistent context architecture combining immutable raw inputs, linked wiki knowledge, line-memory facts, generated outputs, and governed maintenance workflows |
709
- | **[Line Memory](agentic/code/addons/line-memory/)** | Durable append-only facts with retrieval, reinforcement, decay, pruning, and concurrent process safety |
710
- | **[LLM Wiki](agentic/code/addons/llm-wiki/)** | Obsidian-native knowledge base for LLM agents — semantic linking, vault integration |
711
- | **[NLP Prod](agentic/code/addons/nlp-prod/)** | Production NLP pipelines — entity extraction, classification, summarization |
712
- | **[Prose Integration](agentic/code/addons/prose-integration/)** | OpenProse contract grammar integration — declarative service contracts |
713
- | **[Semantic Memory](agentic/code/addons/semantic-memory/)** | Semantic memory kernel — query, capture, lifecycle management for agent memory |
714
- | **[Context Curator](agentic/code/addons/context-curator/)** | Context pre-filtering to remove distractors — production-grade agent reliability |
715
- | **[Verbalized Sampling](agentic/code/addons/verbalized-sampling/)** | Probability distribution prompting — 1.6-2.1x output diversity improvement |
716
- | **[Guided Implementation](agentic/code/addons/guided-implementation/)** | Bounded iteration control for issue-to-code automation |
717
- | **[Skill Factory](agentic/code/addons/skill-factory/)** | Dynamic skill generation and packaging at runtime |
718
- | **[Doc Intelligence](agentic/code/addons/doc-intelligence/)** | Document analysis, PDF extraction, documentation site scraping |
719
- | **[Color Palette](agentic/code/addons/color-palette/)** | WCAG-compliant color palette generation with trend research |
720
- | **[Auto Memory](agentic/code/addons/auto-memory/)** | Automatic memory seed templates for new project context initialization |
721
- | **[Agent Persistence](agentic/code/addons/agent-persistence/)** | Agent state management for session continuity |
722
- | **[AIWG Hooks](agentic/code/addons/aiwg-hooks/)** | Lifecycle event handlers — pre-session, post-write, workflow tracing |
723
- | **[AIWG Utils](agentic/code/addons/aiwg-utils/)** | Core meta-utilities (auto-installed with any framework) |
724
- | **[Droid Bridge](agentic/code/addons/droid-bridge/)** | Factory Droid orchestration — multi-platform agent bridge |
725
- | **[AIWG Fleet](agentic/code/addons/aiwg-fleet/)** | Governed multi-project maintenance with repository discovery, policy-aware planning, approval gates, and auditable execution |
726
- | **[Browser Control](agentic/code/addons/browser-control/)** | Permission-aware browser automation for user-controlled sessions through Playwright MCP |
727
- | **[Twelve-Factor](agentic/code/addons/twelve-factor/)** | Evidence-based application architecture review against Twelve-Factor and modern 12+ Factor criteria |
728
- | **[Star Prompt](agentic/code/addons/star-prompt/)** | Repository star prompt for success celebration |
817
+ AIWG installs reusable context, specialist agents, workflow skills, rules, and
818
+ artifact templates into the AI tools your team already uses. This fragment keeps
819
+ the older README's broad inventory shape while updating claims against the
820
+ current repository. Counts shown in framework rows are source-file counts from
821
+ this working tree; addon rows omit totals because several addons expose
822
+ capabilities through manifests, docs, scripts, or nested skill packages.
823
+
824
+ ### Frameworks
825
+
826
+ | Framework | Source Snapshot | What It Helps You Do |
827
+ |-----------|-----------------|----------------------|
828
+ | **[SDLC Complete](agentic/code/frameworks/sdlc-complete/)** | 100 agents, 116 skills, 217 templates, 39 rules, 12 commands, 8 flows | Run a full software delivery lifecycle from intake through transition with phase gates, planning artifacts, implementation support, test strategy, deployment handoff, and maintenance workflows |
829
+ | **[Forensics Complete](agentic/code/frameworks/forensics-complete/)** | 13 agents, 20 skills, 12 templates, 4 rules | Preserve and analyze incident evidence through scoping, triage, acquisition, log review, persistence hunting, timeline building, IOC extraction, and reporting |
830
+ | **[Media/Marketing Kit](agentic/code/frameworks/media-marketing-kit/)** | 38 agents, 34 skills, 97 templates, 2 flows | Plan, produce, review, publish, and analyze marketing campaigns with reusable briefs, brand/legal gates, channel assets, and performance artifacts |
831
+ | **[Media Curator](agentic/code/frameworks/media-curator/)** | 6 agents, 21 skills | Assess mixed media collections, research sources, acquire approved material, tag metadata, verify integrity, create transcript sidecars, and prepare exports or research handoffs |
832
+ | **[Research Complete](agentic/code/frameworks/research-complete/)** | 8 agents, 41 skills, 16 templates | Turn literature searches and PDFs into reviewable research artifacts: source records, grounded summaries, citation work, GRADE/FAIR-style quality checks, gap notes, and provenance |
833
+ | **[Knowledge Base](agentic/code/frameworks/knowledge-base/)** | 3 skills, 5 templates | Build a linked AI-assisted wiki from loose sources, notes, entities, concepts, comparisons, and synthesis pages without forcing formal literature-review overhead |
834
+ | **[Ops Complete](agentic/code/frameworks/ops-complete/)** | 12 agents, 1 skill, 17 templates, 6 rules | Convert operational procedures into executable runbooks, inventories, incident reports, troubleshooting trees, and extension-backed ops workflows |
835
+ | **[Security Engineering](agentic/code/frameworks/security-engineering/)** | 2 agents, 27 skills, 7 templates, 13 rules | Make applied security decisions for crypto primitives, chains of trust, auth factors, degraded modes, runtime secrets, supply-chain trust, physical threats, and DFIR readiness |
836
+ | **[Validation Complete](agentic/code/frameworks/validation-complete/)** | 1 skill | Add focused validation workflow support where a project needs reviewable checks without adopting a full lifecycle framework |
837
+
838
+ Start with [Install, Connect, and Verify](docs/getting-started/install-connect-verify.md),
839
+ then deploy a framework with `aiwg use <framework>`. The [capability reference](docs/cli/reference.md)
840
+ lists the current framework names accepted by the CLI.
841
+
842
+ ### Addons
843
+
844
+ | Addon | What It Helps You Do |
845
+ |-------|----------------------|
846
+ | **[AIWG Utils](agentic/code/addons/aiwg-utils/)** | Shared rules, discovery helpers, regeneration support, mention tooling, workspace maintenance, and stewardship primitives used across AIWG |
847
+ | **[Agent Loop](agentic/code/addons/agent-loop/)** | Run bounded iterative agent loops with recovery, reflection, completion tracking, and CLI surfaces such as `aiwg ralph` |
848
+ | **[RLM](agentic/code/addons/rlm/)** | Decompose large codebases or document corpora into smaller reviewed slices through recursive planning and subtask execution |
849
+ | **[Composition Engine](agentic/code/addons/composition-engine/)** | Define and validate provider-neutral Flow graph contracts for composed workflows |
850
+ | **[Graph Pattern](agentic/code/addons/graph-pattern/)** | Add an optional graph-oriented profile over AIWG Flow for conditional routes, reducers, and graph validation |
851
+ | **[Orchestration Topology Lab](agentic/code/addons/orchestration-topology-lab/)** | Compare single-agent, bounded-parallel, and planner-worker orchestration topologies using local fixtures and explicit evidence |
852
+ | **[Guided Implementation](agentic/code/addons/guided-implementation/)** | Keep issue-to-code work inside a bounded retry loop with validation after each attempt and structured escalation when needed |
853
+ | **[Daemon](agentic/code/addons/daemon/)** | Run opt-in persistent session support for background tasks, queues, health checks, and scheduler integration |
854
+ | **[Agentic Installer](agentic/code/addons/agentic-installer/)** | Use `setup.aiwg.io/v1` SetupManifest files for reproducible, agent-driven install workflows with recovery paths |
855
+ | **[AIWG Dev](agentic/code/addons/aiwg-dev/)** | Scaffold and validate AIWG source packages, skills, agents, commands, and rules; install explicitly for contributor work |
856
+ | **[Skill Factory](agentic/code/addons/skill-factory/)** | Build, enhance, validate, and package skills through a dedicated skill-authoring workflow |
857
+ | **[AIWG Evals](agentic/code/addons/aiwg-evals/)** | Run agent and workflow evaluation patterns with explicit benchmark inputs and quality scoring |
858
+ | **[Monitorability Red Team](agentic/code/addons/monitorability-red-team/)** | Exercise synthetic local fixtures that expose multi-agent monitoring limits and evidence blind spots |
859
+ | **[Long-Context Bench](agentic/code/addons/long-context-bench/)** | Benchmark compressed skim plus exact recovery against current context baselines |
860
+ | **[Natural-Language Harness](agentic/code/addons/natural-language-harness/)** | Map inspectable natural-language policy documents to deterministic AIWG mechanisms and ablation reports |
861
+ | **[Premortem v2](agentic/code/addons/premortem-v2/)** | Generate, select, and independently verify bounded risk sets before execution |
862
+ | **[Century Readiness](agentic/code/addons/century-readiness/)** | Review long-horizon stewardship, degradation, replacement, evidence, and meaning-preservation risks |
863
+ | **[Dataset Intelligence](agentic/code/addons/dataset-intelligence/)** | Route dataset intake, planning, materialization, traceability, verification, export, synchronization, and retirement through governed workflows |
864
+ | **[Schema Governance](agentic/code/addons/schema-governance/)** | Discover, author, validate, evolve, and normalize schemas across datasets and SDLC artifacts |
865
+ | **[Compound Memory](agentic/code/addons/compound-memory/)** | Govern promotion from raw evidence and session candidates into line memory or linked wiki knowledge with lineage |
866
+ | **[Line Memory](agentic/code/addons/line-memory/)** | Keep a bounded plain-text set of durable project facts with recency retention and reviewed lifecycle operations |
867
+ | **[LLM Wiki](agentic/code/addons/llm-wiki/)** | Maintain a Markdown wiki topology for entities, concepts, sources, comparisons, and syntheses |
868
+ | **[Semantic Memory](agentic/code/addons/semantic-memory/)** | Provide topology-agnostic memory operations for ingest, lint, query/capture, and event logging |
869
+ | **[Auto Memory](agentic/code/addons/auto-memory/)** | Seed Claude Code Automatic Memory files with AIWG-aware testing, debugging, and architecture sections |
870
+ | **[Agent Persistence](agentic/code/addons/agent-persistence/)** | Supply reusable human-in-the-loop gate definitions for destructive actions, overrides, and recovery escalation |
871
+ | **[AIWG Hooks](agentic/code/addons/aiwg-hooks/)** | Provide hook templates for workflow tracing, permissions, session management, context injection, and quality gates |
872
+ | **[AIWG Fleet](agentic/code/addons/aiwg-fleet/)** | Apply quiet-bot, mention-only participation, and small-plan cost-discipline policies across multi-project fleets |
873
+ | **[Browser Control](agentic/code/addons/browser-control/)** | Drive a user-authorized Chromium-derived browser through Playwright MCP with allow-list and audit boundaries |
874
+ | **[Droid Bridge](agentic/code/addons/droid-bridge/)** | Bridge Claude Code to Factory Droid for batch operations and automated fixes through MCP |
875
+ | **[MCP/UAT Toolkit](agentic/code/addons/uat-mcp/)** | Generate, execute, and report user-acceptance tests against MCP tool surfaces |
876
+ | **[Civic Action](agentic/code/addons/civic-action/)** | Prepare evidence-bound civic research, public-records planning, meeting review, local-resource profiles, corrections, and publication review |
877
+ | **[Network Analysis](agentic/code/addons/network-analysis/)** | Governed saved-PCAP/PCAPNG analysis with bounded TShark recipes, cited packet evidence, and optional local Termshark review |
878
+ | **[Testing Quality](agentic/code/addons/testing-quality/)** | Assess test conformance, normalize suites with reversible plans, and add TDD, mutation, flaky-test, and factory workflows |
879
+ | **[Writing Quality](agentic/code/addons/writing-quality/)** | Review editorial quality, author requirements, and voice consistency without treating heuristic scores as authorship proof |
880
+ | **[Voice Framework](agentic/code/addons/voice-framework/)** | Define, analyze, blend, and apply reusable writing voice profiles and runtime-selectable output modes |
881
+ | **[Color Palette](agentic/code/addons/color-palette/)** | Generate and review accessible color palettes using color theory, trend research, and WCAG checks |
882
+ | **[Doc Intelligence](agentic/code/addons/doc-intelligence/)** | Scrape, extract, split, audit, and synchronize documentation sources |
883
+ | **[Prose Integration](agentic/code/addons/prose-integration/)** | Detect, read, validate, wire, and run OpenProse contract programs in supported AIWG sessions |
884
+ | **[NLP Prod](agentic/code/addons/nlp-prod/)** | Design and productionize LLM inference pipelines with eval-first, pattern-guided workflow support |
885
+ | **[Context Curator](agentic/code/addons/context-curator/)** | Filter distractors and curate context packs for agent work where irrelevant material can derail results |
886
+ | **[Twelve-Factor](agentic/code/addons/twelve-factor/)** | Review or design applications against Twelve-Factor and modern cloud-native criteria |
887
+ | **[Verbalized Sampling](agentic/code/addons/verbalized-sampling/)** | Apply and evaluate verbalized probability-distribution prompting for output diversity experiments |
888
+ | **[Star Prompt](agentic/code/addons/star-prompt/)** | Offer a tasteful repository-star prompt after successful command completion |
889
+
890
+ Addon details live in each source directory and, where public docs exist, under
891
+ `docs/addons/`. Use [Key Addons](docs/getting-started/key-addons.md) for a
892
+ guided end-user selection path.
729
893
 
730
894
  ---
731
895
 
732
- ### Agents (188)
896
+ ### Agents
733
897
 
734
- Specialized AI personas deployed to your platform with defined tools, responsibilities, and operating rhythms.
898
+ Specialized AI personas deploy to your platform with defined responsibilities,
899
+ tools, and operating rhythms. The exact inventory changes as frameworks evolve,
900
+ so this README keeps durable groupings and examples instead of relying on one
901
+ global total.
735
902
 
736
- #### SDLC Agents (90)
903
+ #### SDLC Agents
737
904
 
738
- | Domain | Agents | Examples |
739
- |--------|--------|---------|
740
- | **Testing & Quality** | 11 | Test Engineer, Test Architect, Mutation Analyst, Regression Analyst, Laziness Detector, Reliability Engineer |
741
- | **Security & Compliance** | 9 | Security Auditor, Security Architect, Compliance Checker, Privacy Officer, Citation Verifier |
742
- | **Architecture & Design** | 12 | Architecture Designer, API Designer, Cloud Architect, System Analyst, Product Designer, Decision Matrix Expert |
743
- | **DevOps & Cloud** | 8 | AWS Specialist, Azure Specialist, GCP Specialist, Kubernetes Expert, DevOps Engineer, Multi-Cloud Strategist |
744
- | **Backend & Data** | 10 | Django Expert, Spring Boot Expert, Data Engineer, Database Optimizer, Software Implementer, Incident Responder |
745
- | **Frontend & Mobile** | 6 | React Expert, Frontend Specialist, Mobile Developer, Accessibility Specialist, UX Lead |
746
- | **AI/ML & Performance** | 5 | AI/ML Engineer, Performance Engineer, Cost Optimizer, Metrics Analyst |
747
- | **Code Quality** | 11 | Code Reviewer, Debugger, Dead Code Analyzer, Technical Debt Analyst, Legacy Modernizer |
748
- | **Documentation** | 7 | Technical Writer, Documentation Synthesizer, Documentation Archivist, Context Librarian |
749
- | **Requirements & Planning** | 7 | Requirements Analyst, Requirements Reviewer, Intake Coordinator, RACI Expert |
750
- | **Agent/Tool Smiths** | 9 | AgentSmith, CommandSmith, MCPSmith, SkillSmith, ToolSmith |
751
- | **Governance & Meta** | 3 | Executive Orchestrator, Recovery Orchestrator, Migration Planner |
905
+ | Domain | Examples |
906
+ |--------|----------|
907
+ | **Testing & Quality** | Test Engineer, Test Architect, Mutation Analyst, Regression Analyst, Reliability Engineer |
908
+ | **Security & Compliance** | Security Auditor, Security Architect, Compliance Checker, Privacy Officer, Citation Verifier |
909
+ | **Architecture & Design** | Architecture Designer, API Designer, Cloud Architect, System Analyst, Product Designer, Decision Matrix Expert |
910
+ | **DevOps & Cloud** | AWS Specialist, Azure Specialist, GCP Specialist, Kubernetes Expert, DevOps Engineer, Multi-Cloud Strategist |
911
+ | **Backend & Data** | Django Expert, Spring Boot Expert, Data Engineer, Database Optimizer, Software Implementer, Incident Responder |
912
+ | **Frontend & Mobile** | React Expert, Frontend Specialist, Mobile Developer, Accessibility Specialist, UX Lead |
913
+ | **AI/ML & Performance** | AI/ML Engineer, Performance Engineer, Cost Optimizer, Metrics Analyst |
914
+ | **Code Quality** | Code Reviewer, Debugger, Dead Code Analyzer, Technical Debt Analyst, Legacy Modernizer |
915
+ | **Documentation** | Technical Writer, Documentation Synthesizer, Documentation Archivist, Context Librarian |
916
+ | **Requirements & Planning** | Requirements Analyst, Requirements Reviewer, Intake Coordinator, RACI Expert |
917
+ | **Agent/Tool Smiths** | AgentSmith, CommandSmith, MCPSmith, SkillSmith, ToolSmith |
918
+ | **Governance & Meta** | Executive Orchestrator, Recovery Orchestrator, Migration Planner |
752
919
 
753
- #### Forensics Agents (13)
920
+ #### Forensics Agents
754
921
 
755
922
  | Agent | What It Does |
756
923
  |-------|-------------|
757
- | Forensics Orchestrator | Coordinates full investigation lifecycle from scoping through reporting |
758
- | Triage Agent | Quick volatile data capture following RFC 3227 volatility order |
759
- | Acquisition Agent | Evidence collection with chain of custody and SHA-256 hash verification |
760
- | Log Analyst | Auth.log, syslog, journal, and application log analysis for brute force, privilege escalation, lateral movement |
761
- | Persistence Hunter | Sweeps cron, systemd, SSH keys, LD_PRELOAD, PAM modules, kernel modules — maps to MITRE ATT&CK |
762
- | Container Analyst | Docker, containerd, Kubernetes forensics — privilege escalation, container escapes, eBPF monitoring |
763
- | Network Analyst | Connection state, DNS, traffic patterns — beaconing, C2, data exfiltration detection |
764
- | Memory Analyst | Volatility 3 memory forensics — process analysis, rootkit detection, credential extraction |
765
- | Cloud Analyst | AWS/Azure/GCP audit logs, IAM review, network flows, API activity anomaly detection |
766
- | Timeline Builder | Multi-source event correlation — chronological incident timelines with attribution |
767
- | IOC Analyst | IOC extraction, enrichment, STIX 2.1 formatting — actionable IOC register |
768
- | Recon Agent | Target reconnaissance — system topology, services, users, network baselines |
769
- | Reporting Agent | Structured forensic reports — executive summary, technical findings, timeline, remediation |
770
-
771
- #### Marketing Agents (37)
772
-
773
- | Domain | Agents |
774
- |--------|--------|
924
+ | Forensics Orchestrator | Coordinates investigation scope, evidence handling, analysis, and reporting |
925
+ | Triage Agent | Captures volatile data following evidence-priority guidance |
926
+ | Acquisition Agent | Collects evidence with chain-of-custody and hash verification |
927
+ | Log Analyst | Reviews auth, syslog, journal, and application logs for suspicious activity |
928
+ | Persistence Hunter | Checks cron, systemd, SSH keys, LD_PRELOAD, PAM modules, and kernel-module indicators |
929
+ | Container Analyst | Reviews Docker, containerd, and Kubernetes evidence |
930
+ | Network Analyst | Reviews connection state, DNS, beaconing, and exfiltration indicators |
931
+ | Memory Analyst | Supports Volatility-style memory forensics workflows |
932
+ | Cloud Analyst | Reviews AWS, Azure, and GCP audit trails and IAM posture |
933
+ | Timeline Builder | Correlates events into chronological incident timelines |
934
+ | IOC Analyst | Extracts and formats indicators for downstream response |
935
+ | Recon Agent | Builds a target baseline for authorized investigation |
936
+ | Reporting Agent | Produces structured executive and technical investigation reports |
937
+
938
+ #### Marketing Agents
939
+
940
+ | Domain | Examples |
941
+ |--------|----------|
775
942
  | **Strategy** | Campaign Strategist, Brand Guardian, Positioning Specialist, Market Researcher, Content Strategist, Channel Strategist |
776
943
  | **Creation** | Copywriter, Content Writer, Email Marketer, Social Media Specialist, SEO Specialist, Graphic Designer, Art Director |
777
944
  | **Management** | Campaign Orchestrator, Production Coordinator, Traffic Manager, Asset Manager, Workflow Coordinator |
778
945
  | **Analytics** | Marketing Analyst, Data Analyst, Attribution Specialist, Reporting Specialist, Budget Planner |
779
946
  | **Communications** | PR Specialist, Crisis Communications, Corporate Communications, Internal Communications, Media Relations |
780
947
 
781
- #### Research Agents (8)
948
+ #### Other Framework Agents
782
949
 
783
- Discovery Agent, Acquisition Agent, Documentation Agent, Citation Agent, Quality Agent, Archival Agent, Provenance Agent, Workflow Agent
784
-
785
- #### Media Curator Agents (6)
786
-
787
- Discography Analyst, Source Discoverer, Acquisition Manager, Quality Assessor, Metadata Curator, Completeness Tracker
950
+ Research uses discovery, acquisition, documentation, citation, quality,
951
+ archival, provenance, and workflow roles. Media Curator uses discography/source,
952
+ acquisition, quality, metadata, and completeness roles. Ops Complete adds
953
+ runbook execution and inventory roles. Security Engineering adds security
954
+ specialists for applied security decisions and supply-chain review.
788
955
 
789
956
  ---
790
957
 
791
- ### Rules (35)
958
+ ### Rules
792
959
 
793
- Enforcement patterns that prevent common AI failure modes. Rules deploy automatically with their framework.
960
+ Rules are durable guardrails that deploy with the frameworks or addons that own
961
+ them. They prevent common agent failure modes and define review boundaries.
794
962
 
795
- #### Core Rules (10) — Always Active
963
+ #### Core Rules
796
964
 
797
965
  | Rule | Severity | What It Enforces |
798
966
  |------|----------|-----------------|
799
- | `no-attribution` | CRITICAL | AI tools are tools — never add attribution to commits, PRs, docs, or code |
800
- | `token-security` | CRITICAL | Never hard-code tokens; use heredoc pattern for scoped lifetime; file permissions 600 |
801
- | `versioning` | CRITICAL | CalVer YYYY.M.PATCH with NO leading zeros; npm rejects leading zeros |
802
- | `citation-policy` | CRITICAL | Never fabricate citations, DOIs, or URLs; only cite verified sources; GRADE-appropriate hedging |
803
- | `anti-laziness` | HIGH | Never delete tests to pass, skip tests, remove features, or weaken assertions; escalate after 3 failures |
804
- | `executable-feedback` | HIGH | Execute tests before returning code; track execution history; max 3 retries with root cause analysis |
805
- | `failure-mitigation` | HIGH | Detect and recover from 6 LLM failure archetypes: hallucination, context loss, instruction drift, safety, technical, consistency |
806
- | `research-before-decision` | HIGH | Research codebase before acting: IDENTIFY → SEARCH → EXTRACT → REASON → ACT → VERIFY |
807
- | `instruction-comprehension` | HIGH | Fully parse all instructions before acting; track multi-part requests to completion |
808
- | `subagent-scoping` | HIGH | One focused task per subagent; <20% context budget; no delegation chains deeper than 2 levels |
809
-
810
- #### SDLC Rules (34) — Active with Framework
811
-
812
- Actionable feedback, mention wiring, HITL gates, agent fallback, provenance tracking, TAO loop, reproducibility validation, SDLC orchestration, agent-friendly code, agent generation guardrails, artifact discovery, HITL patterns, human gate display, thought protocol, reasoning sections, few-shot examples, best output selection, reproducibility, progressive disclosure, conversable agent interface, auto-reply chains, criticality panel sizing, qualified references.
813
-
814
- #### Research Rules (2) — Active with Research
815
-
816
- Research metadata (FAIR-compliant YAML frontmatter), index generation (auto-generated INDEX.md per FAIR F4).
967
+ | `no-attribution` | CRITICAL | AI tools are tools; do not add AI attribution to commits, PRs, docs, or code |
968
+ | `token-security` | CRITICAL | Keep tokens and secrets out of source; use scoped lifetime and restricted file permissions |
969
+ | `versioning` | CRITICAL | Use the repository's CalVer release format consistently |
970
+ | `citation-policy` | CRITICAL | Do not fabricate citations, DOIs, URLs, or research claims |
971
+ | `anti-laziness` | HIGH | Do not delete tests, skip required checks, remove features, or weaken assertions to pass |
972
+ | `executable-feedback` | HIGH | Run appropriate validation before returning implementation work |
973
+ | `failure-mitigation` | HIGH | Detect and recover from hallucination, context loss, instruction drift, safety, technical, and consistency failures |
974
+ | `research-before-decision` | HIGH | Inspect the codebase and docs before making technical decisions |
975
+ | `instruction-comprehension` | HIGH | Parse prohibitions, requirements, and preferences before acting |
976
+ | `subagent-scoping` | HIGH | Keep delegated tasks focused and bounded when delegation is used |
977
+
978
+ #### Domain Rules
979
+
980
+ | Domain | Examples |
981
+ |--------|----------|
982
+ | **SDLC** | HITL gates, provenance tracking, artifact discovery, phase gates, reproducibility validation, agent-friendly code, fallback, review, and handoff rules |
983
+ | **Forensics** | Evidence integrity, chain of custody, forensic reporting, and authorized investigation boundaries |
984
+ | **Security Engineering** | Cryptographic decision boundaries, runtime secret hygiene, supply-chain trust, physical-access threat modeling, and DFIR readiness handoff |
985
+ | **Ops** | Ops safety, executable runbook format, evidence governance, issue tracking, and cross-repo reference rules |
986
+ | **Civic Action** | Human authority, citation, publication, public-source, privacy, and anti-targeting boundaries |
987
+ | **Addon Rules** | Browser authorization, dataset boundaries, agentic installer safety, voice/output behavior, and hook discipline |
817
988
 
818
989
  ---
819
990
 
820
- ### Skills (128)
821
-
822
- Natural language workflow triggers. Say "what's the project status?" and the `project-awareness` skill activates.
823
-
824
- | Category | Skills | Examples |
825
- |----------|--------|---------|
826
- | **Regression Testing** | 12 | regression-check, regression-baseline, regression-bisect, regression-performance, regression-api-contract, regression-cicd-hooks, regression-learning |
827
- | **Voice & Writing** | 6 | voice-create, voice-analyze, voice-apply, voice-blend, ai-pattern-detection, brand-compliance |
828
- | **Testing & Quality** | 8 | auto-test-execution, test-coverage, test-sync, mutation-test, flaky-detect, flaky-fix, tdd-enforce, qa-protocol |
829
- | **Forensics & Security** | 8 | linux-forensics, memory-forensics, cloud-forensics, container-forensics, sigma-hunting, log-analysis, ioc-extraction, supply-chain-forensics |
830
- | **SDLC & Workflow** | 10 | sdlc-accelerate, sdlc-reports, gate-evaluation, approval-workflow, iteration-control, risk-cycle, parallel-dispatch, decision-support |
831
- | **Documentation** | 6 | doc-sync, doc-scraper, doc-splitter, llms-txt-support, pdf-extractor, source-unifier |
832
- | **Artifacts & Traceability** | 6 | artifact-orchestration, artifact-metadata, artifact-lookup, traceability-check, claims-validator, citation-guard |
833
- | **Research** | 2 | grade-on-ingest, auto-provenance |
834
- | **Infrastructure** | 5 | config-validator, template-engine, code-chunker, decompose-file, workspace-health |
835
- | **Iteration** | 4 | agent-loop, issue-driven-ralph, cross-task-learner, reflection-injection |
836
- | **Other** | 19 | performance-digest, competitive-intel, audience-synthesis, skill-builder, skill-enhancer, skill-packager, quality-checker, nl-router, tot-exploration, and more |
991
+ ### Skills
992
+
993
+ Skills are natural-language workflows. A user describes an outcome, the agent
994
+ discovers the relevant skill, loads its instructions, and applies its protocol.
995
+ The current repo contains a large and changing skill surface, so this section
996
+ keeps durable categories and examples.
997
+
998
+ | Category | Examples |
999
+ |----------|---------|
1000
+ | **Capability discovery and setup** | `aiwg-utils-quickref`, `steward`, `aiwg-status`, `aiwg-doctor`, `use`, provider regeneration |
1001
+ | **SDLC and delivery** | `intake-wizard`, `sdlc-accelerate`, gate evaluation, delivery-track flows, deployment, guided implementation |
1002
+ | **Testing and quality** | `test-conformance`, `test-normalize`, `test-platform-research`, TDD, mutation, flaky-test review, factory generation |
1003
+ | **Security and forensics** | supply-chain hardening, auth-factor design, degraded-mode review, DFIR readiness, log analysis, IOC extraction |
1004
+ | **Research and knowledge** | source acquisition, paper induction, GRADE checks, citation work, wiki ingest, synthesis, knowledge-base health |
1005
+ | **Marketing and content** | campaign intake, creative brief, brand compliance, social strategy, email campaigns, performance digests |
1006
+ | **Media curation** | source discovery, acquisition planning, transcript sidecars, metadata tagging, quality filtering, archive verification |
1007
+ | **Datasets and schemas** | dataset intake, source assessment, capability recommendation, plan review, ingest, trace, verify, export, retire |
1008
+ | **Memory and persistence** | line-memory operations, compound-memory review, semantic-memory capture/query, llm-wiki topology |
1009
+ | **Operations and automation** | runbook execution, ops verification, activity logs, hooks, daemon sessions, schedule support |
1010
+ | **Authoring and development** | skill creation, addon/framework scaffolding, validation, schema governance, doc synchronization |
837
1011
 
838
1012
  ---
839
1013
 
@@ -841,9 +1015,11 @@ Natural language workflow triggers. Say "what's the project status?" and the `pr
841
1015
 
842
1016
  ### SDLC Complete — Full Software Development Lifecycle
843
1017
 
844
- The SDLC framework implements a phase-gated development lifecycle with 90 specialized agents, 34 enforcement rules, and 170+ artifact templates. Natural language commands drive phase transitions with automated quality gates.
1018
+ The SDLC framework implements a phase-gated development lifecycle with
1019
+ specialized agents, enforcement rules, and artifact templates. Natural-language
1020
+ requests drive phase transitions with reviewable quality gates.
845
1021
 
846
- ```
1022
+ ```text
847
1023
  ┌──────────┐ ┌─────────────┐ ┌──────────────┐ ┌────────────┐ ┌────────────┐
848
1024
  │ CONCEPT │───▶│ INCEPTION │───▶│ ELABORATION │───▶│CONSTRUCTION│───▶│ TRANSITION │
849
1025
  │ │ │ │ │ │ │ │ │ │
@@ -854,7 +1030,7 @@ The SDLC framework implements a phase-gated development lifecycle with 90 specia
854
1030
  └──────────┘ └──────┬──────┘ └──────┬───────┘ └─────┬──────┘ └────────────┘
855
1031
  │ │ │
856
1032
  ┌──▼──┐ ┌──▼──┐ ┌──▼──┐
857
- │ LOM │ │ ABM │ │ IOC │ ← Quality Gates
1033
+ │ LOM │ │ ABM │ │ IOC │
858
1034
  │Gate │ │Gate │ │Gate │
859
1035
  └─────┘ └─────┘ └─────┘
860
1036
 
@@ -862,36 +1038,36 @@ The SDLC framework implements a phase-gated development lifecycle with 90 specia
862
1038
  IOC = Initial Operational Capability
863
1039
  ```
864
1040
 
865
- **SDLC Flow Commands (24):**
1041
+ **SDLC Flow Commands:**
866
1042
 
867
1043
  | Command | Phase | What It Does |
868
1044
  |---------|-------|-------------|
869
- | `/intake-wizard` | Concept | Generate project intake form from natural language description |
870
- | `/intake-start` | Concept→Inception | Validate intake, kick off with agent assignments |
871
- | `/intake-from-codebase` | Concept | Scan existing codebase, generate intake from analysis |
872
- | `/flow-concept-to-inception` | Concept→Inception | Phase transition with intake validation and vision alignment |
873
- | `/flow-inception-to-elaboration` | Inception→Elaboration | Architecture baselining and risk retirement |
874
- | `/flow-elaboration-to-construction` | Elaboration→Construction | Iteration planning, team scaling, full-scale development |
875
- | `/flow-construction-to-transition` | Construction→Transition | IOC validation, production deployment, operational handover |
876
- | `/flow-discovery-track` | Any | Prepare validated requirements one iteration ahead of delivery |
877
- | `/flow-delivery-track` | Any | Test-driven development, quality gates, iteration assessment |
878
- | `/flow-iteration-dual-track` | Any | Synchronized Discovery + Delivery workflows |
879
- | `/flow-deploy-to-production` | Transition | Strategy selection, validation, automated rollback, regression gates |
880
- | `/flow-incident-response` | Operations | Triage, escalation, resolution, post-incident review (ITIL) |
881
- | `/flow-security-review-cycle` | Any | Continuous security validation, threat modeling, vulnerability management |
882
- | `/flow-performance-optimization` | Any | Baseline, bottleneck ID, optimization, load testing, SLO validation |
883
- | `/flow-retrospective-cycle` | Any | Structured feedback, improvement tracking, action items |
884
- | `/flow-change-control` | Any | Baseline management, impact assessment, CCB review, communication |
885
- | `/flow-risk-management-cycle` | Any | Continuous risk identification, assessment, tracking, retirement |
886
- | `/flow-compliance-validation` | Any | Requirements mapping, audit evidence, gap analysis, attestation |
887
- | `/flow-knowledge-transfer` | Transition | Assessment, documentation, shadowing, validation, handover |
888
- | `/flow-team-onboarding` | Any | Pre-boarding, training, buddy assignment, 30/60/90 day check-ins |
889
- | `/flow-hypercare-monitoring` | Transition | 24/7 support, SLO tracking, rapid issue response |
890
- | `/flow-gate-check` | Any | Multi-agent phase gate validation with comprehensive reporting |
891
- | `/flow-handoff-checklist` | Any | Handoff validation between phases and tracks |
892
- | `/flow-guided-implementation` | Construction | Bounded iteration with issue-to-code automation |
893
-
894
- **SDLC Accelerate — Idea to Construction-Ready in One Command:**
1045
+ | `/intake-wizard` | Concept | Generate project intake from a natural-language description |
1046
+ | `/intake-start` | Concept -> Inception | Validate intake and begin agent assignments |
1047
+ | `/intake-from-codebase` | Concept | Scan an existing codebase and generate intake from analysis |
1048
+ | `/flow-concept-to-inception` | Concept -> Inception | Transition with intake validation and vision alignment |
1049
+ | `/flow-inception-to-elaboration` | Inception -> Elaboration | Baseline architecture and retire major risks |
1050
+ | `/flow-elaboration-to-construction` | Elaboration -> Construction | Prepare iteration planning, scale delivery, and begin implementation |
1051
+ | `/flow-construction-to-transition` | Construction -> Transition | Validate IOC, deployment readiness, and operational handoff |
1052
+ | `/flow-discovery-track` | Any | Prepare validated requirements ahead of delivery |
1053
+ | `/flow-delivery-track` | Any | Run test-driven delivery with quality gates |
1054
+ | `/flow-iteration-dual-track` | Any | Coordinate discovery and delivery tracks |
1055
+ | `/flow-deploy-to-production` | Transition | Select deployment strategy, validate, and prepare rollback/regression checks |
1056
+ | `/flow-incident-response` | Operations | Triage, resolve, and review incidents |
1057
+ | `/flow-security-review-cycle` | Any | Run continuous security validation and threat review |
1058
+ | `/flow-performance-optimization` | Any | Baseline, identify bottlenecks, optimize, and validate SLOs |
1059
+ | `/flow-retrospective-cycle` | Any | Capture feedback and track improvement actions |
1060
+ | `/flow-change-control` | Any | Assess impact, coordinate review, and manage communication |
1061
+ | `/flow-risk-management-cycle` | Any | Identify, assess, track, and retire risks |
1062
+ | `/flow-compliance-validation` | Any | Map requirements, collect evidence, and identify gaps |
1063
+ | `/flow-knowledge-transfer` | Transition | Prepare documentation, shadowing, validation, and handover |
1064
+ | `/flow-team-onboarding` | Any | Structure onboarding, training, buddy support, and follow-up |
1065
+ | `/flow-hypercare-monitoring` | Transition | Track early-life support, SLOs, and rapid-response items |
1066
+ | `/flow-gate-check` | Any | Run multi-agent phase-gate validation |
1067
+ | `/flow-handoff-checklist` | Any | Validate handoff between phases and tracks |
1068
+ | `/flow-guided-implementation` | Construction | Run bounded issue-to-code iteration with validation and escalation |
1069
+
1070
+ **SDLC Accelerate — from idea to reviewed planning artifacts:**
895
1071
 
896
1072
  ```bash
897
1073
  # From a description
@@ -904,11 +1080,13 @@ aiwg sdlc-accelerate --from-codebase .
904
1080
  aiwg sdlc-accelerate --resume
905
1081
  ```
906
1082
 
907
- Generates intake form, vision document, use cases, architecture baseline, risk register, test strategy, and deployment plan — all with human approval gates between phases.
1083
+ It can generate intake, vision, use cases, architecture baseline, risk register,
1084
+ test strategy, and deployment planning artifacts with human review between
1085
+ major phases.
908
1086
 
909
1087
  **Dual-Track Iteration Model:**
910
1088
 
911
- ```
1089
+ ```text
912
1090
  ┌─────────────────────────────────────────────────┐
913
1091
  │ ITERATION N │
914
1092
  │ │
@@ -931,21 +1109,23 @@ Generates intake form, vision document, use cases, architecture baseline, risk r
931
1109
  └─────────────────────────────────────────────────┘
932
1110
  ```
933
1111
 
934
- **Metrics & Quality Tracking:**
1112
+ **Metrics and Quality Tracking:**
935
1113
 
936
1114
  | Metric Category | Metrics Tracked |
937
1115
  |-----------------|-----------------|
938
- | **DORA** (4) | Deployment Frequency, Lead Time, Change Failure Rate, MTTR |
939
- | **Velocity** (3) | Story Points, Cycle Time, Throughput |
940
- | **Flow** (3) | WIP Limits, Flow Efficiency, Blocked Items |
941
- | **Quality** (13) | Test Coverage (4), Defect Metrics (4), Code Quality (3), Technical Debt (2) |
942
- | **Operational** (16) | SLO/SLI (5), Infrastructure (4), Incidents (4), Cost (3) |
1116
+ | **DORA** | Deployment frequency, lead time, change failure rate, MTTR |
1117
+ | **Velocity** | Story points, cycle time, throughput |
1118
+ | **Flow** | WIP limits, flow efficiency, blocked items |
1119
+ | **Quality** | Test coverage, defect metrics, code quality, technical debt |
1120
+ | **Operational** | SLO/SLI, infrastructure, incidents, cost |
943
1121
 
944
- ### Forensics Complete — Digital Forensics & Incident Response
1122
+ ### Forensics Complete — Digital Forensics and Incident Response
945
1123
 
946
- Full DFIR investigation workflow following NIST SP 800-86, with MITRE ATT&CK mapping and Sigma rule hunting.
1124
+ Forensics Complete supports authorized DFIR work following NIST SP 800-86-style
1125
+ evidence handling, MITRE ATT&CK mapping, Sigma hunting, timeline construction,
1126
+ and structured reporting.
947
1127
 
948
- ```
1128
+ ```text
949
1129
  ┌──────────┐ ┌──────────┐ ┌────────────┐ ┌──────────┐ ┌──────────┐
950
1130
  │ SCOPE │───▶│ TRIAGE │───▶│ ACQUIRE │───▶│ ANALYZE │───▶│ REPORT │
951
1131
  │ │ │ │ │ │ │ │ │ │
@@ -962,29 +1142,16 @@ Full DFIR investigation workflow following NIST SP 800-86, with MITRE ATT&CK map
962
1142
  **Investigation Commands:**
963
1143
 
964
1144
  ```bash
965
- /forensics-profile # Build target system profile via SSH
966
- /forensics-triage # Quick triage following RFC 3227 volatility order
967
- /forensics-acquire # Evidence acquisition with chain of custody
968
- /forensics-investigate # Full multi-agent investigation workflow
969
- /forensics-timeline # Build correlated event timeline
970
- /forensics-hunt # Threat hunt using Sigma rules
971
- /forensics-ioc # Extract and enrich IOCs
972
- /forensics-report # Generate forensic investigation report
973
- /forensics-status # Show investigation dashboard
974
- ```
975
-
976
- **Bundled Sigma Rules (8):**
977
-
978
- | Rule | What It Detects |
979
- |------|----------------|
980
- | SSH Brute Force | Repeated failed SSH authentication attempts |
981
- | Unauthorized SUID | Unexpected SUID/SGID binaries |
982
- | LD_PRELOAD Rootkit | Library injection via LD_PRELOAD |
983
- | Cron Persistence | Unauthorized crontab modifications |
984
- | Kernel Module Load | Suspicious kernel module insertion |
985
- | PAM Backdoor | PAM configuration tampering |
986
- | SSH Key Injection | Unauthorized authorized_keys modifications |
987
- | Systemd Persistence | Suspicious systemd unit creation |
1145
+ /forensics-profile
1146
+ /forensics-triage
1147
+ /forensics-acquire
1148
+ /forensics-investigate
1149
+ /forensics-timeline
1150
+ /forensics-hunt
1151
+ /forensics-ioc
1152
+ /forensics-report
1153
+ /forensics-status
1154
+ ```
988
1155
 
989
1156
  **Supported Evidence Sources:**
990
1157
 
@@ -992,49 +1159,70 @@ Full DFIR investigation workflow following NIST SP 800-86, with MITRE ATT&CK map
992
1159
  |--------|-------|----------|
993
1160
  | Auth logs | Log Analyst | Brute force, privilege escalation, lateral movement |
994
1161
  | Syslog / journal | Log Analyst | System events, service anomalies |
995
- | Network connections | Network Analyst | C2 beaconing, data exfiltration, DNS tunneling |
996
- | Docker/containerd | Container Analyst | Container escapes, image tampering, eBPF monitoring |
997
- | Memory dumps | Memory Analyst | Process injection, rootkits, credential extraction |
998
- | AWS/Azure/GCP | Cloud Analyst | API anomalies, IAM abuse, network flow analysis |
1162
+ | Network connections | Network Analyst | C2 beaconing, exfiltration, DNS tunneling |
1163
+ | Docker/containerd | Container Analyst | Container escape, image tampering, runtime evidence |
1164
+ | Memory dumps | Memory Analyst | Process analysis, rootkits, credential artifacts |
1165
+ | AWS/Azure/GCP | Cloud Analyst | API anomalies, IAM abuse, network-flow evidence |
999
1166
  | File system | Persistence Hunter | Cron, systemd, SSH keys, PAM, kernel modules |
1000
1167
 
1001
1168
  ### Media/Marketing Kit — Campaign Lifecycle
1002
1169
 
1003
- ```
1170
+ Media/Marketing Kit treats campaign work as a lifecycle with artifacts and
1171
+ review gates, so strategy, content, legal/brand review, publication planning,
1172
+ and performance analysis remain inspectable.
1173
+
1174
+ ```text
1004
1175
  ┌──────────┐ ┌──────────┐ ┌──────────┐ ┌──────────┐ ┌──────────┐
1005
1176
  │ STRATEGY │───▶│ CREATION │───▶│ REVIEW │───▶│ PUBLISH │───▶│ ANALYZE │
1006
1177
  │ │ │ │ │ │ │ │ │ │
1007
1178
  │ Research │ │ Copy │ │ Brand │ │ Schedule │ │ KPIs │
1008
1179
  │ Audience │ │ Design │ │ Legal │ │ Channels │ │ Reports │
1009
- │ Strategy │ │ Content │ │ Quality │ │ Launch │ │ ROI │
1180
+ │ Strategy │ │ Content │ │ Quality │ │ Launch │ │ Learnings│
1010
1181
  └──────────┘ └──────────┘ └──────────┘ └──────────┘ └──────────┘
1011
1182
  ```
1012
1183
 
1013
- 37 agents across strategy, creation, management, analytics, and communications. 87+ templates covering campaign intake, brand guidelines, content briefs, social playbooks, email sequences, PR kits, and analytics dashboards.
1184
+ | Discipline | Example Artifacts |
1185
+ |------------|-------------------|
1186
+ | Strategy | Campaign intake, positioning, messaging, audience profile, channel plan |
1187
+ | Creation | Blog drafts, social posts, email sequences, creative briefs, media kits |
1188
+ | Review | Brand compliance, legal clearance, accessibility review, claim substantiation |
1189
+ | Publication | Launch checklist, schedule, channel handoff, go-live readiness |
1190
+ | Analysis | KPI report, performance digest, retrospective, optimization plan |
1191
+
1192
+ ### Media Curator — Archive Management
1014
1193
 
1015
- ### Media Curator — Intelligent Archive Management
1194
+ Media Curator helps assess, acquire, organize, verify, transcribe, and export
1195
+ media collections. It starts with assessment and planning so unknown or mixed
1196
+ media is routed before downloads or metadata rewrites.
1016
1197
 
1017
1198
  ```bash
1018
1199
  # Full curation pipeline
1019
1200
  /curate "Pink Floyd"
1020
1201
 
1021
1202
  # Step by step
1022
- /analyze-artist "Pink Floyd" # Identify eras, catalog structure
1023
- /find-sources "Pink Floyd" "DSOTM" # Discover across YouTube, Archive.org, Bandcamp
1024
- /acquire # Download with format selection
1025
- /transcribe-media /path/to/media.wav # Create timestamped transcript sidecars
1026
- /tag-collection # Apply metadata, embed artwork, rename
1027
- /check-completeness # Gap analysis against canonical discography
1028
- /assemble "Pink Floyd live 1973" # Build thematic compilations
1029
- /export --format plex # Export to Plex, Jellyfin, MPD, or archival
1030
- /verify-archive # SHA-256 integrity verification
1031
- ```
1032
-
1033
- Quality tiers: Tier 1 (Official/Lossless) → Tier 2 (High Quality) → Tier 3 (Acceptable) → Tier 4 (Avoid). Transcript sidecars preserve source hashes, transcript hashes, timestamps, and optional speaker labels for review and future research handoff. Standards: ID3v2.4, Vorbis Comments, MusicBrainz, PREMIS 3.0, W3C PROV-O.
1203
+ /analyze-artist "Pink Floyd"
1204
+ /find-sources "Pink Floyd" "DSOTM"
1205
+ /acquire
1206
+ /transcribe-media /path/to/media.wav
1207
+ /tag-collection
1208
+ /check-completeness
1209
+ /assemble "Pink Floyd live 1973"
1210
+ /export --format plex
1211
+ /verify-archive
1212
+ ```
1213
+
1214
+ Quality tiers help reviewers choose what to keep. Transcript sidecars preserve
1215
+ source hashes, transcript hashes, timestamps, and optional speaker labels for
1216
+ review and later research handoff. Common standards include ID3v2.4, Vorbis
1217
+ Comments, MusicBrainz, PREMIS 3.0, and W3C PROV-O.
1034
1218
 
1035
1219
  ### Research Complete — Academic Research Pipeline
1036
1220
 
1037
- ```
1221
+ Research Complete turns search results and PDFs into source-grounded,
1222
+ reviewable research artifacts with persistent identifiers, quality checks, and
1223
+ provenance.
1224
+
1225
+ ```text
1038
1226
  ┌──────────┐ ┌──────────┐ ┌──────────┐ ┌──────────┐
1039
1227
  │ DISCOVER │───▶│ ACQUIRE │───▶│ DOCUMENT │───▶│ ARCHIVE │
1040
1228
  │ │ │ │ │ │ │ │
@@ -1045,20 +1233,92 @@ Quality tiers: Tier 1 (Official/Lossless) → Tier 2 (High Quality) → Tier 3 (
1045
1233
  └──────────┘ └──────────┘ └──────────┘ └──────────┘
1046
1234
  ```
1047
1235
 
1048
- 8-stage pipeline: Discovery → Acquisition → Documentation → Citation → Quality Assessment → Synthesis → Gap Analysis → Archival. Persistent REF-XXX identifiers. GRADE scoring (HIGH/MODERATE/LOW/VERY LOW). Unpaywall integration for open access papers.
1236
+ Pipeline stages: Discovery -> Acquisition -> Documentation -> Citation ->
1237
+ Quality Assessment -> Synthesis -> Gap Analysis -> Archival. The framework uses
1238
+ `REF-XXX` identifiers, GRADE-style evidence quality labels, FAIR-style checks,
1239
+ and Unpaywall lookup for open-access discovery. It flags unsupported claims for
1240
+ review instead of promising error-free summaries.
1241
+
1242
+ ### Knowledge Base — Linked Project Wiki
1243
+
1244
+ Knowledge Base is for open-ended knowledge accumulation where the taxonomy
1245
+ emerges over time. It uses entity, concept, source, comparison, and synthesis
1246
+ pages so future sessions can find what is known, what is missing, and how ideas
1247
+ connect.
1248
+
1249
+ | Page Type | Purpose |
1250
+ |-----------|---------|
1251
+ | Entity | A person, company, tool, place, system, or other named thing |
1252
+ | Concept | A technique, pattern, framework, or idea |
1253
+ | Source | The evidence layer for claims and summaries |
1254
+ | Comparison | A decision aid for tools, approaches, vendors, or options |
1255
+ | Synthesis | A higher-level claim produced by combining multiple sources |
1256
+
1257
+ ### Ops Complete — Executable Operations
1258
+
1259
+ Ops Complete gives operational procedures a structured envelope: inventory,
1260
+ capabilities, playbooks, gates, targets, schedules, pipelines, and extensions.
1261
+ It is useful when procedures must be idempotent, verifiable, and evidence-aware.
1262
+
1263
+ ```yaml
1264
+ apiVersion: ops.aiwg.io/v1
1265
+ kind: OpsPlaybook
1266
+ metadata:
1267
+ name: deploy-auth-stack
1268
+ namespace: production
1269
+ spec:
1270
+ # Desired state
1271
+ status:
1272
+ # Observed state written by the executor
1273
+ ```
1274
+
1275
+ Extensions add domain-specific ops support for systems, IT, development
1276
+ infrastructure, and streaming workflows. See [Ops Complete overview](docs/frameworks/ops-complete/overview.md)
1277
+ and [Ops evidence governance](https://github.com/jmagly/aiwg/blob/main/docs/ops-evidence-governance.md).
1278
+
1279
+ ### Security Engineering — Applied Security Decisions
1280
+
1281
+ Security Engineering complements SDLC and forensics by focusing on security
1282
+ decisions that need explicit assumptions and reviewable tradeoffs.
1283
+
1284
+ | Area | Example Use |
1285
+ |------|-------------|
1286
+ | Cryptographic primitives | Choose AEAD, KDF, hashing, randomness, and signing patterns for a concrete workload |
1287
+ | Chain of trust | Map trust anchors, update paths, verification points, and failure modes |
1288
+ | Authentication factors | Decide factor mix, enrollment, recovery, lockout, and degraded-mode behavior |
1289
+ | Runtime secret hygiene | Review secret storage, process boundaries, rotation, logging, and local development exposure |
1290
+ | Supply-chain trust | Review dependency sources, lifecycle scripts, release provenance, SBOMs, and signed artifacts |
1291
+ | Physical-access threats | Model device seizure, kiosk, lab, field, and hostile-local-user conditions |
1292
+ | DFIR readiness | Prepare evidence handoff points before an incident occurs |
1293
+
1294
+ Install with:
1295
+
1296
+ ```bash
1297
+ aiwg use security-engineering
1298
+ ```
1299
+
1300
+ ### Validation Complete — Focused Validation
1301
+
1302
+ Validation Complete provides a small validation workflow surface for teams that
1303
+ need structured review without adopting a broader lifecycle. Use it where a
1304
+ project already has its own process but wants AIWG-style validation gates and
1305
+ reports.
1049
1306
 
1050
1307
  ---
1051
1308
 
1052
1309
  ## Voice Framework — Content Voice Consistency
1053
1310
 
1054
- 4 built-in voice profiles with create, analyze, blend, and apply skills:
1311
+ Voice Framework defines reusable writing profiles and output modes that can be
1312
+ applied across docs, release notes, campaigns, reports, and internal guidance.
1313
+ It describes the desired voice directly rather than relying only on banned word
1314
+ lists.
1055
1315
 
1056
1316
  | Profile | When to Use | Characteristics |
1057
1317
  |---------|-------------|----------------|
1058
- | `technical-authority` | API docs, architecture guides | Precise terminology, confident assertions, specific metrics |
1059
- | `friendly-explainer` | Tutorials, onboarding | Accessible language, analogies, encouragement |
1060
- | `executive-brief` | Status reports, proposals | Bottom-line-first, quantified impact, action-oriented |
1061
- | `casual-conversational` | Blog posts, social media | Natural rhythm, opinions, varied structure |
1318
+ | `technical-authority` | API docs, architecture guides | Precise terminology, direct claims, concrete examples |
1319
+ | `friendly-explainer` | Tutorials, onboarding | Accessible language, patient sequencing, light warmth |
1320
+ | `executive-brief` | Status reports, proposals | Decision-oriented summaries, concise evidence, clear next steps |
1321
+ | `casual-conversational` | Blog posts, social media | Natural rhythm, opinion-forward phrasing, varied structure |
1062
1322
 
1063
1323
  ```bash
1064
1324
  # Apply a voice to content
@@ -1072,626 +1332,806 @@ Quality tiers: Tier 1 (Official/Lossless) → Tier 2 (High Quality) → Tier 3 (
1072
1332
 
1073
1333
  # Blend two voices
1074
1334
  /voice-blend technical-authority casual-conversational --ratio 70:30
1075
-
1076
- # Detect AI patterns and suggest authentic alternatives
1077
- /ai-pattern-detection docs/generated-content.md
1078
1335
  ```
1079
1336
 
1080
- ---
1337
+ See [Voice Framework overview](docs/addons/voice-framework/overview.md) and
1338
+ [Voice Framework quickstart](docs/addons/voice-framework/quickstart.md).
1081
1339
 
1082
1340
  ## MCP Server — Model Context Protocol Integration
1083
1341
 
1084
- AIWG includes a built-in MCP server for tool-based AI workflow integration:
1342
+ AIWG can expose its project context, discovery catalog, and governed workflows through a Model Context Protocol
1343
+ server. This lets MCP-capable tools call AIWG without learning the repository layout or memorizing provider-specific
1344
+ file locations.
1345
+
1346
+ The MCP server is useful when you want an assistant to ask AIWG questions such as “what capabilities are available for
1347
+ release planning?”, “show the SDLC quickstart”, or “run this governed workflow and return the evidence artifact.” The
1348
+ base server keeps a small default tool surface; larger toolsets can be enabled explicitly for teams that want richer
1349
+ orchestration, mission, dataset, or framework operations.
1085
1350
 
1086
1351
  ```bash
1087
- # Start MCP server
1352
+ # Run the local MCP server
1088
1353
  aiwg mcp serve
1089
1354
 
1090
- # Install into Claude Desktop
1091
- aiwg mcp install claude
1355
+ # Enable additional toolsets for a richer host integration
1356
+ aiwg mcp serve --toolsets=flows,missions,catalog
1357
+
1358
+ # Use an environment variable when the host launches the server command
1359
+ AIWG_MCP_TOOLSETS=flows,missions,catalog aiwg mcp serve
1092
1360
 
1093
- # Show capabilities
1361
+ # Inspect server metadata and supported install targets
1094
1362
  aiwg mcp info
1095
1363
  ```
1096
1364
 
1097
- The MCP server exposes AIWG's artifact management, workflow execution, and project health capabilities as tools that any MCP-compatible AI platform can invoke programmatically.
1365
+ Provider installation depends on the host. AIWG can write MCP configuration for supported targets where the provider
1366
+ has a stable local MCP config format; other providers use the same server command in their own UI or settings.
1098
1367
 
1099
- ---
1368
+ ```bash
1369
+ # Install AIWG MCP config for a supported local target
1370
+ aiwg mcp install claude
1371
+ aiwg mcp install cursor
1372
+ aiwg mcp install codex
1373
+ ```
1374
+
1375
+ MCP integration does not make every AIWG operation model-backed. Catalog reads, status checks, link resolution, and
1376
+ local evidence inspection are ordinary local operations. Workflows that ask an assistant to reason, draft, call
1377
+ another provider, or continue a Ralph loop may use model calls depending on the connected host and selected provider.
1378
+
1379
+ See also: [MCP server documentation](docs/mcp/README.md), [MCP capability
1380
+ audit](docs/integrations/mcp-capability-audit.md), and [cross-platform
1381
+ overview](docs/integrations/cross-platform-overview.md).
1100
1382
 
1101
1383
  ## Agent Evaluation Framework
1102
1384
 
1103
- Test agent quality with archetype resistance testing based on Roig (2025) failure patterns:
1385
+ AIWG treats agents, skills, commands, and rules as reviewable project assets. The evaluation workflow is designed to
1386
+ answer concrete questions before you rely on a capability in a live project:
1387
+
1388
+ - Does the capability declare the right trigger conditions and boundaries?
1389
+ - Does it cite the files, schemas, or rules it depends on?
1390
+ - Does it produce artifacts that another provider can inspect?
1391
+ - Does it fail safely when prerequisites are missing?
1392
+ - Does it preserve evidence for audit, handoff, or regression review?
1393
+
1394
+ For direct CLI checks, use the catalog, metadata validation, skill linting, and evidence commands that are available
1395
+ in the current CLI surface.
1104
1396
 
1105
1397
  ```bash
1106
- # Evaluate a specific agent
1107
- /eval-agent security-auditor
1398
+ # Find relevant evaluation or review capabilities
1399
+ aiwg discover "agent evaluation" --limit 5
1400
+ aiwg discover "skill lint evidence" --limit 5
1401
+
1402
+ # Inspect a selected capability before applying it
1403
+ aiwg show skill aiwg-doctor
1404
+ aiwg show skill context-firewall
1108
1405
 
1109
- # Test archetype resistance
1110
- /eval-agent test-engineer --category archetype
1406
+ # Validate local metadata and skill packaging
1407
+ aiwg validate-metadata
1408
+ aiwg skill-lint
1111
1409
 
1112
- # Run performance benchmarks
1113
- /eval-agent code-reviewer --category performance
1410
+ # Capture and verify evidence bundles when a workflow supports them
1411
+ aiwg evidence export --help
1412
+ aiwg evidence verify --help
1114
1413
  ```
1115
1414
 
1116
- **Test Categories:**
1415
+ A practical evaluation usually starts with a small task and a pass/fail criterion. For example:
1117
1416
 
1118
- | Category | Tests | What It Validates |
1119
- |----------|-------|-------------------|
1120
- | **Archetype** | 4 | Grounding (hallucination resistance), Substitution (scope adherence), Distractor (context noise), Recovery (failure handling) |
1121
- | **Performance** | 3 | Latency, token efficiency, parallel execution capability |
1122
- | **Quality** | 3 | Output format compliance, correct tool usage, scope adherence |
1417
+ > Evaluate the release-note drafting capability against this repository. Use only committed changelog entries and
1418
+ merged PR metadata. The result is acceptable if every claim links to a source artifact and uncited claims are listed
1419
+ separately.
1123
1420
 
1124
- Target score: >=85% per agent. Results include passed/failed breakdown with evidence.
1125
-
1126
- ---
1421
+ That prompt-led path is intentional. AIWG can route the work through provider-native tools, but the acceptance
1422
+ criterion remains explicit and reviewable. Avoid treating any score, pass rate, or runtime as guaranteed across models
1423
+ or providers; those values depend on the selected model, available tools, project size, and the evidence the workflow
1424
+ can inspect.
1127
1425
 
1128
1426
  ## Bidirectional Traceability — @-Mention System
1129
1427
 
1130
- Link requirements to architecture to code to tests with semantic @-mentions:
1428
+ AIWG uses lightweight `@` references to connect instructions, generated artifacts, source files, evidence, and
1429
+ follow-up work. The goal is traceability across providers: a model can move from a rule to the code it governs, from a
1430
+ generated report to the source data behind it, or from an issue to the artifact that closed it.
1131
1431
 
1132
1432
  ```markdown
1133
- <!-- In a use case document -->
1134
- This use case @implements(UC-001) the authentication flow
1135
- described in @architecture(SAD-section-3.2).
1136
-
1137
- <!-- In a test file -->
1138
- // @tests(UC-001) @depends(auth-service)
1139
- describe('authentication flow', () => { ... });
1433
+ <!-- In an agent or skill file -->
1434
+ @src/auth/middleware.ts
1435
+ @docs/security/authentication.md
1436
+ @.aiwg/evidence/release-2026-09-07.json
1140
1437
  ```
1141
1438
 
1142
- **Mention Commands:**
1439
+ Traceability matters most when a workflow crosses boundaries. An SDLC intake can reference the use case it created. A
1440
+ context-firewall review can reference the baseline it approved. A dataset query can reference the source, ingest plan,
1441
+ checkpoint, and verification record instead of relying on conversation memory.
1442
+
1443
+ Provider-facing commands and skills may expose mention helpers such as mention wiring, validation, linting, or
1444
+ reporting. Because providers package commands differently, the portable entry point is discovery:
1143
1445
 
1144
1446
  ```bash
1145
- /mention-wire # Analyze codebase and inject @-mentions for traceability
1146
- /mention-validate # Validate all @-mentions resolve to existing files
1147
- /mention-report # Generate traceability report
1148
- /mention-lint # Lint @-mentions for style consistency
1149
- /mention-conventions # Display naming conventions and placement rules
1447
+ aiwg discover "mention validate" --limit 5
1448
+ aiwg discover "traceability report" --limit 5
1449
+ aiwg show skill context-firewall
1150
1450
  ```
1151
1451
 
1152
- Relationship qualifiers: `@implements`, `@tests`, `@depends`, `@derives-from`, `@blocked-by`, `@supersedes`. Enables queries like "what implements UC-001?" and "what tests cover the auth module?"
1452
+ When deployed into a provider that supports slash commands, the same work is often available as a prompt command, for example:
1153
1453
 
1154
- ---
1454
+ ```text
1455
+ /mention-validate --target docs
1456
+ /mention-report --scope .aiwg/reports
1457
+ ```
1458
+
1459
+ Use root-relative references in public documentation so links work from the README. Use provider-specific absolute
1460
+ paths only inside generated provider files where that provider requires them.
1155
1461
 
1156
1462
  ## Configuration & Customization
1157
1463
 
1158
- ### Workspace Structure
1159
-
1160
- ```
1161
- your-project/
1162
- ├── .aiwg/ # SDLC artifacts (persistent project memory)
1163
- │ ├── intake/ # Project intake forms
1164
- │ ├── requirements/ # Use cases, user stories, NFRs
1165
- │ ├── architecture/ # SAD, ADRs, system diagrams
1166
- │ ├── planning/ # Phase plans, iteration plans
1167
- │ ├── risks/ # Risk register, mitigations
1168
- │ ├── testing/ # Test strategy, plans, results
1169
- │ ├── security/ # Threat models, security gates
1170
- │ ├── deployment/ # Deployment plans, runbooks
1171
- │ ├── reports/ # Generated status reports
1172
- │ ├── ralph/ # In-session agent loop state
1173
- │ ├── ralph-external/ # External Ralph crash-resilient state
1174
- │ ├── research/ # Research corpus and findings
1175
- │ ├── forensics/ # Investigation artifacts
1176
- │ ├── working/ # Temporary files (safe to delete)
1177
- │ └── frameworks/ # Installed framework registry
1178
- │ └── registry.json
1179
- ├── .claude/ # Claude Code deployment
1180
- │ ├── agents/ # 162 agent definitions
1181
- │ ├── commands/ # Slash commands
1182
- │ ├── skills/ # 86 skill definitions
1183
- │ └── rules/ # RULES-INDEX.md + on-demand full rules
1184
- ├── .github/ # GitHub Copilot deployment
1185
- ├── .cursor/ # Cursor deployment
1186
- ├── .warp/ # Warp Terminal deployment
1187
- └── CLAUDE.md # Project instructions (auto-generated)
1464
+ AIWG separates project context from provider packaging. The project keeps canonical context and generated artifacts
1465
+ under the workspace, then deploys provider-specific adapters for Claude, Codex, Cursor, Windsurf, Warp, OpenCode,
1466
+ OpenClaw, OpenHuman, Hermes, DeepSeek Harness, Copilot, Devin, Factory, Oh My Pi, Pi Coding Agent, Antigravity, and
1467
+ the generic fallback.
1468
+
1469
+ The primary files are:
1470
+
1471
+ ```text
1472
+ WORKSPACE.md # Project/operator context read by providers
1473
+ AIWG.md # AIWG discovery and routing guide
1474
+ AGENTS.md # Provider bootstrap for Codex and other AGENTS.md readers
1475
+ .aiwg/ # Canonical AIWG config, generated context, evidence, reports
1476
+ .aiwg/aiwg.config # Workspace configuration and provider deployment state
1477
+ .aiwg/index/ # Searchable artifact and capability indexes when generated
1478
+ .aiwg/reports/ # Audits, sync reports, doctor reports, workflow outputs
1479
+ .aiwg/sessions/ # Optional local session catalog data
1480
+ .aiwg/datasets/ # Optional dataset plans, manifests, lineage, and exports
1481
+ ```
1482
+
1483
+ Provider directories are generated from the same canonical context. Their exact shape depends on the provider:
1484
+
1485
+ ```text
1486
+ .claude/ # Claude Code skills, commands, hooks, settings
1487
+ .codex/ or ~/.codex/ # Codex prompts and global configuration where applicable
1488
+ .agents/ # Cross-provider agents and skills used by Codex/Antigravity/OMP
1489
+ .cursor/ # Cursor rules and skills
1490
+ .github/ # GitHub Copilot prompts, instructions, and agents
1491
+ .warp/ # Warp skills and compatibility assets
1492
+ .omp/ # Oh My Pi native agents, prompts, rules, and bootstrap
1188
1493
  ```
1189
1494
 
1190
1495
  ### Creating Custom Extensions
1191
1496
 
1497
+ Use the current scaffolding commands for new AIWG assets. The older scaffold commands remain available in some
1498
+ workspaces for compatibility, but the `new-*` and `add-*` commands are the clearer path for new work.
1499
+
1500
+ ```bash
1501
+ # Create a new bundle, extension, addon, framework, or provider adapter
1502
+ aiwg new-bundle my-bundle
1503
+ aiwg new-extension my-extension
1504
+ aiwg new-addon my-addon
1505
+ aiwg new-framework my-framework
1506
+ aiwg new-provider my-provider
1507
+
1508
+ # Add individual provider-facing assets
1509
+ aiwg add-agent release-reviewer --framework sdlc
1510
+ aiwg add-command release-checklist --framework sdlc
1511
+ aiwg add-skill release-notes --framework sdlc
1512
+
1513
+ # Validate metadata before sharing or deploying
1514
+ aiwg validate-metadata
1515
+ ```
1516
+
1517
+ A custom extension should define the smallest durable contract needed by the workflow: triggers, inputs, outputs,
1518
+ evidence, and provider packaging. Keep model-specific phrasing in provider assets. Keep project policy, schemas, and
1519
+ reusable workflow contracts in `.aiwg` or extension source so multiple providers can share them.
1520
+
1521
+ ### Capability Discovery — `aiwg discover` + `aiwg show`
1522
+
1523
+ `aiwg discover` searches the installed AIWG capability catalog. It is the recommended entry point when you know the
1524
+ task but not the command, skill, agent, or framework name.
1525
+
1192
1526
  ```bash
1193
- # Add a custom agent
1194
- aiwg add-agent my-domain-expert
1527
+ # Find capabilities by plain-language intent
1528
+ aiwg discover "deploy production" --limit 5
1529
+ aiwg discover "dataset lineage" --type skill --limit 5
1530
+ aiwg discover "SDLC intake requirements" --limit 8
1531
+
1532
+ # Inspect the selected capability before running or asking a provider to use it
1533
+ aiwg show skill aiwg-status
1534
+ aiwg show skill dataset-intelligence
1535
+ aiwg show skill rlm-prep
1536
+ ```
1195
1537
 
1196
- # Add a custom command
1197
- aiwg add-command my-workflow
1538
+ Discovery is also useful for documentation. Instead of hard-coding every command in a README, link to the relevant
1539
+ quickstart and show one or two representative commands. The catalog can change as frameworks and addons are installed,
1540
+ while the task language stays stable.
1198
1541
 
1199
- # Add a custom skill
1200
- aiwg add-skill my-capability
1542
+ ### Artifact Index — `aiwg index`
1201
1543
 
1202
- # Scaffold a complete addon
1203
- aiwg scaffold-addon my-addon
1544
+ The artifact index makes generated work easier to find, verify, and reuse. It indexes reports, generated context,
1545
+ evidence, and other AIWG-managed files into a searchable local catalog.
1204
1546
 
1205
- # Scaffold a complete framework
1206
- aiwg scaffold-framework my-framework
1547
+ ```bash
1548
+ # Build or refresh the local artifact index
1549
+ aiwg index
1207
1550
 
1208
- # Validate all extension metadata
1209
- aiwg validate-metadata
1551
+ # Inspect status after deployment or refresh
1552
+ aiwg status --probe
1553
+ aiwg doctor
1210
1554
  ```
1211
1555
 
1212
- ### Capability Discovery — `aiwg discover` + `aiwg show`
1556
+ A provider can then answer questions such as “find the latest context-firewall report” or “show the SDLC artifact that
1557
+ introduced this acceptance criterion” without scanning the whole repository manually.
1558
+
1559
+ ### Doc Sync — Bidirectional Documentation
1213
1560
 
1214
- The headline operator surface for finding and reading AIWG capabilities. Most AIWG skills (~455 of 480+) are **not loaded into your platform's flat skill listing** — they stay at `$AIWG_ROOT` and are reached on demand through `aiwg discover` (find) and `aiwg show` (fetch). The kernel set on disk is small on purpose: 9 framework quickrefs + 16 self-maintenance and discovery skills = 25 skills, within supported provider listing budgets.
1561
+ Doc sync is for keeping code and documentation aligned under review. It can audit mismatches, propose updates, and
1562
+ write reports before code or documentation changes are accepted.
1215
1563
 
1216
1564
  ```bash
1217
- # Find a skill by capability
1218
- aiwg discover "deploy production" # → flow-deploy-to-production
1219
- aiwg discover "create intake" # → intake-* family
1220
- aiwg discover "audit security" --type skill --limit 5
1221
- aiwg discover "<phrase>" --format json # stable ids for sub-agents, no paths
1222
- aiwg discover "<phrase>" --format json --compact
1565
+ # Audit both directions without writing changes
1566
+ aiwg doc-sync full --dry-run --scope docs
1223
1567
 
1224
- # Fetch the full body of a specific artifact (companion to discover)
1225
- aiwg show skill aiwg:skill:6f1477d99813ca8d
1226
- aiwg show skill flow-deploy-to-production
1227
- aiwg show agent aiwg-steward
1228
- aiwg show command discover
1229
- aiwg show rule no-attribution
1568
+ # Propose documentation updates from code changes
1569
+ aiwg doc-sync code-to-docs --scope src --guidance "Update quickstarts only"
1230
1570
 
1231
- # Inspect Fortemi metadata and resolved paths when needed
1232
- aiwg show metadata aiwg:skill:6f1477d99813ca8d --json
1571
+ # Propose code TODOs or implementation tasks from documentation requirements
1572
+ aiwg doc-sync docs-to-code --scope docs --interactive
1233
1573
  ```
1234
1574
 
1235
- The kernel quickrefs ship **curated, validated discovery phrases per capability domain** — phrases tested against the live scorer to surface the right top-3 candidates. The self-maintenance and discovery set (including `steward`, `aiwg-doctor`, `aiwg-refresh`, `aiwg-status`, `aiwg-help`, and `use`) stays loaded so the agent retains repair surfaces even when discovery itself is broken. See [`docs/discovery-and-kernel-skills.md`](docs/discovery-and-kernel-skills.md) for the full best-practices guide, ASCII flow diagrams, and verification steps.
1575
+ Doc sync writes reports under `.aiwg/working/` and `.aiwg/reports/` when configured. Treat those reports as review
1576
+ artifacts. Do not assume doc sync can prove semantic equivalence between code and prose; it identifies
1577
+ inconsistencies, stale examples, missing links, and candidate updates for human or provider review.
1236
1578
 
1237
- ### Artifact Index — `aiwg index`
1579
+ ### Reproducibility Validation
1580
+
1581
+ AIWG’s reproducibility features focus on explicit inputs, evidence records, deterministic modes where available, and
1582
+ reviewable outputs. They do not guarantee identical model text across providers or runs.
1238
1583
 
1239
1584
  ```bash
1240
- # Build searchable artifact index
1241
- aiwg index build
1242
- aiwg index build --force --verbose
1585
+ # Put local execution in a stricter mode for workflows that honor it
1586
+ aiwg execution-mode strict --seed 12345
1243
1587
 
1244
- # Search artifacts by keyword
1245
- aiwg index query "authentication" --json
1588
+ # Export and verify evidence when workflows emit evidence bundles
1589
+ aiwg evidence export --help
1590
+ aiwg evidence verify --help
1246
1591
 
1247
- # Show dependency graph for an artifact
1248
- aiwg index deps .aiwg/requirements/UC-001.md --json
1592
+ # Verify workspace health and generated provider context
1593
+ aiwg verify --help
1594
+ aiwg doctor
1595
+ ```
1249
1596
 
1250
- # Index statistics
1251
- aiwg index stats --json
1597
+ For workflows that expose checkpointing, snapshots, or replay through installed skills, start with discovery so the
1598
+ current workspace selects the correct implementation:
1599
+
1600
+ ```bash
1601
+ aiwg discover "create checkpoint" --limit 5
1602
+ aiwg discover "replay evidence" --limit 5
1252
1603
  ```
1253
1604
 
1254
- The index supports multiple graphs: project graph (`.aiwg/` artifacts), codebase graph (`src/` / `test/` / `tools/`), and framework graph (`agentic/code/` + `docs/`).
1605
+ The practical standard is repeatability of inputs, citations, commands, and artifacts. Exact model wording should be
1606
+ treated as a generated output, not as the source of truth.
1255
1607
 
1256
- ### Doc Sync — Bidirectional Documentation
1608
+ ### Session Catalog
1609
+
1610
+ The session catalog is an optional local feature for importing, searching, and promoting useful provider conversation
1611
+ history. It is designed for controlled handoff and audit. It should be enabled intentionally because it can include
1612
+ sensitive prompts, local paths, and project context.
1257
1613
 
1258
1614
  ```bash
1259
- # Audit doc drift (dry run)
1260
- aiwg doc-sync code-to-docs --dry-run
1615
+ # Install the SQLite feature before using the session catalog
1616
+ aiwg features install sqlite
1261
1617
 
1262
- # Sync docs to match code
1263
- aiwg doc-sync code-to-docs
1618
+ # Discover importable sessions for this workspace without changing state
1619
+ aiwg sessions discover --workspace "$PWD" --dry-run
1264
1620
 
1265
- # Bidirectional reconciliation
1266
- aiwg doc-sync full --interactive
1621
+ # Import discovered sessions after review
1622
+ aiwg sessions import-discovered --workspace "$PWD" --confirm
1623
+
1624
+ # Inspect, search, and audit imported sessions
1625
+ aiwg sessions list
1626
+ aiwg sessions timeline
1627
+ aiwg sessions search "release blocker"
1628
+ aiwg sessions doctor
1267
1629
  ```
1268
1630
 
1269
- ### Reproducibility Validation
1631
+ Imported sessions can be tagged, extracted into reusable notes, reviewed for promotion, or audited for provenance.
1632
+ Keep private-provider roots and shared history locations explicit in configuration; do not assume another provider’s
1633
+ global history is safe to import by default.
1634
+
1635
+ See [session history setup](docs/getting-started/session-history.md) and [sessions CLI](docs/sessions/cli.md).
1636
+
1637
+ ### Dataset Intelligence
1638
+
1639
+ Dataset intelligence gives AIWG a governed path for local files, directories, CSV/JSONL sources, and approved HTTP
1640
+ sources. The dataset router carries stable source, plan, checkpoint, lineage, verification, and export references
1641
+ between phases.
1270
1642
 
1271
1643
  ```bash
1272
- # Show/set execution mode (strict = temperature 0, fixed seed)
1273
- aiwg execution-mode
1644
+ # Register a source from a JSON descriptor
1645
+ aiwg dataset source --file dataset-source.json --json
1274
1646
 
1275
- # Create execution snapshot
1276
- aiwg snapshot
1647
+ # Check and preview before ingestion
1648
+ aiwg dataset check source:docs --json
1649
+ aiwg dataset preview source:docs --count 5 --offline
1277
1650
 
1278
- # Create workflow checkpoint
1279
- aiwg checkpoint
1651
+ # Create and approve an ingest plan
1652
+ aiwg dataset plan --file dataset-plan.json --json
1653
+ aiwg dataset ingest plan:docs-index \
1654
+ --digest sha256:<approved-plan-digest> \
1655
+ --idempotency-key docs-index-2026-09-07
1280
1656
 
1281
- # Validate workflow reproducibility
1282
- aiwg reproducibility-validate
1657
+ # Inspect and use the resulting dataset
1658
+ aiwg dataset status dataset:docs-index
1659
+ aiwg dataset verify dataset:docs-index
1660
+ aiwg dataset query dataset:docs-index "Which quickstart explains Codex setup?"
1661
+ aiwg dataset lineage dataset:docs-index
1662
+ aiwg dataset export dataset:docs-index --json
1283
1663
  ```
1284
1664
 
1285
- Thresholds: compliance audit (100%), security scan (100%), test generation (95%).
1665
+ Local adapters are constrained by configured roots. HTTP adapters are deny-by-default and require explicit hosts.
1666
+ Indexes are derived artifacts; the canonical record is the source descriptor, approved plan, ingest run, and evidence
1667
+ trail.
1286
1668
 
1287
- ---
1669
+ See [dataset intelligence quickstart](docs/addons/dataset-intelligence/quickstart.md), [dataset
1670
+ overview](docs/addons/dataset-intelligence/overview.md), and [source adapters](docs/dataset/source-adapters.md).
1288
1671
 
1289
1672
  ## Issue-Driven Development
1290
1673
 
1291
- AIWG integrates with issue trackers for 2-way human-AI collaboration:
1674
+ AIWG supports local issue planning and governed handoff to external issue trackers. The local issue CLI stores records
1675
+ under `.aiwg/issues/`, which makes issues reviewable even when a project does not have GitHub, Gitea, Jira, or another
1676
+ tracker connected.
1292
1677
 
1293
1678
  ```bash
1294
- # Create issues from any backend (Gitea, GitHub, Jira, Linear, local files)
1295
- /issue-create "Implement OAuth2 flow" --labels "feature,auth"
1679
+ # Initialize a local issue store for this workspace
1680
+ aiwg issue init --prefix APP
1681
+
1682
+ # Draft a new local issue
1683
+ aiwg issue plan \
1684
+ --title "Implement OAuth2 callback validation" \
1685
+ --body "Add state validation, token exchange error handling, and tests."
1296
1686
 
1297
- # List and filter issues
1298
- /issue-list --state open --labels "priority:high"
1687
+ # Review and update issues locally
1688
+ aiwg issue list --status open --label auth --limit 20
1689
+ aiwg issue show APP-0001 --comments last:10
1690
+ aiwg issue comment APP-0001 --body "Validated the callback edge cases."
1691
+ aiwg issue close APP-0001 --reason "Implemented and tested."
1692
+ ```
1299
1693
 
1300
- # Drive an issue with agent loop — posts status to issue thread
1301
- /issue-driven-ralph 42
1694
+ External tracker import/export is explicit. Use it when you need traceability between local AIWG records and a remote
1695
+ system, and keep snapshots or live connector settings under review.
1302
1696
 
1303
- # Auto-sync issues from commits and artifacts
1304
- /issue-sync
1697
+ ```bash
1698
+ # Import a tracker snapshot into the local issue store
1699
+ aiwg issue import --from github --snapshot-file issues-snapshot.json
1305
1700
 
1306
- # Close with comprehensive summary and verification
1307
- /issue-close 42
1701
+ # Export a local issue payload for a tracker
1702
+ aiwg issue export APP-0001 --to gitea --out APP-0001.gitea.json
1703
+
1704
+ # Inspect conflicts when reconciling local and external state
1705
+ aiwg issue sync conflicts APP-0001 --snapshot-file issue-APP-0001.json
1308
1706
  ```
1309
1707
 
1310
- The `/address-issues` command orchestrates issue-thread-driven agent loops with automatic progress posting and human feedback incorporation at each cycle.
1708
+ For agent-assisted repair work, ask for the issue outcome directly and include the acceptance checks. In providers
1709
+ with deployed prompt commands, `/address-issues` can route the work through the configured workflow.
1311
1710
 
1312
- ---
1711
+ ```text
1712
+ /address-issues APP-0001 APP-0002 --checks "npm test && npm run lint"
1713
+ ```
1714
+
1715
+ External issue systems are not a default side effect of `aiwg issue`. They require configured connectors, snapshots,
1716
+ or explicit export/import commands. See [local issue integration](docs/local-issues.md) and [filing
1717
+ issues](docs/contributing/filing-issues.md).
1313
1718
 
1314
1719
  ## Daemon Mode & Messaging Integration
1315
1720
 
1316
- ### Daemon Mode
1721
+ AIWG’s automation layer is for long-running coordination, not for hiding work from review. The safe default is local,
1722
+ explicit execution with visible status and evidence. Daemon, messaging, and mission-control setups should declare
1723
+ their trigger source, operator identity, workspace, budget limits, and completion criteria.
1724
+
1725
+ The base CLI exposes current orchestration commands through Ralph and mission control. Messaging bridges and chat bots
1726
+ are advanced deployments described in the daemon and messaging docs; they require external service configuration and
1727
+ should not be assumed to exist in a fresh checkout.
1317
1728
 
1318
1729
  ```bash
1319
- # Background file watching, cron scheduling, IPC
1320
- aiwg daemon start
1730
+ # Start a managed mission-control session
1731
+ aiwg mc start --name "release follow-up" --max-missions 3
1732
+
1733
+ # Dispatch bounded work with an explicit completion criterion
1734
+ aiwg mc dispatch <session-id> \
1735
+ "Fix the failing auth tests" \
1736
+ --completion "npm test -- auth passes"
1737
+
1738
+ # Inspect and control running work
1739
+ aiwg mc run <session-id>
1740
+ aiwg mc status <session-id>
1741
+ aiwg mc watch <session-id>
1742
+ aiwg mc pause <session-id>
1743
+ aiwg mc resume <session-id>
1744
+ aiwg mc stop <session-id>
1321
1745
  ```
1322
1746
 
1323
- See [Daemon Guide](docs/daemon-guide.md) for background agent orchestration.
1747
+ For provider messaging, document the concrete external channel and approval boundary. A Slack, Discord, Telegram, or
1748
+ webhook bridge should make it clear who can enqueue work, where logs are stored, and which operations require human
1749
+ approval before writing to external systems.
1750
+
1751
+ See [daemon guide](docs/daemon-guide.md), [messaging guide](docs/messaging-guide.md), and [Mission Control](docs/addons/ralph/quickstart.md).
1752
+
1753
+ ## See It In Action
1324
1754
 
1325
- ### Messaging Integration
1755
+ The fastest way to use AIWG is to ask for the first useful task, request a concrete deliverable, and name the success
1756
+ check. Commands help when you know the exact workflow; plain-language task prompts are better when AIWG should choose
1757
+ the relevant skill or provider surface.
1326
1758
 
1327
- Bidirectional Slack, Discord, and Telegram bots for remote agent control:
1759
+ ### SDLC workflow from idea to implementation
1328
1760
 
1329
1761
  ```bash
1330
- # Connect to messaging platforms
1331
- aiwg messaging connect slack
1332
- aiwg messaging connect discord
1333
- aiwg messaging connect telegram
1762
+ # Discover the right SDLC entry point
1763
+ aiwg discover "SDLC intake requirements architecture" --limit 5
1764
+
1765
+ # Run the accelerator when you want AIWG to scaffold the SDLC work plan
1766
+ aiwg sdlc-accelerate "AI-powered code review tool" \
1767
+ --success "requirements, architecture notes, and first implementation task are generated"
1334
1768
  ```
1335
1769
 
1336
- See [Messaging Guide](docs/messaging-guide.md) for setup and configuration.
1770
+ Provider prompt:
1337
1771
 
1338
- ---
1772
+ ```text
1773
+ Use the SDLC framework to turn “AI-powered code review tool” into requirements, architecture decisions, a first implementation task, and acceptance checks. Stop with links to the generated artifacts.
1774
+ ```
1339
1775
 
1340
- ## See It In Action
1776
+ ### Long-running implementation loop
1341
1777
 
1342
1778
  ```bash
1343
- # Generate project intake from natural language
1344
- /intake-wizard "Build customer portal with real-time chat"
1779
+ aiwg ralph "Fix all failing tests in the auth package" \
1780
+ --completion "npm test -- auth passes" \
1781
+ --max-iterations 5 \
1782
+ --max-wall-clock-minutes 45
1345
1783
 
1346
- # Accelerate from idea to construction-ready
1347
- /sdlc-accelerate "AI-powered code review tool"
1784
+ aiwg ralph-status
1785
+ aiwg ralph-resume <loop-id>
1786
+ aiwg ralph-abort <loop-id>
1787
+ ```
1348
1788
 
1349
- # Phase transition with automated gate check
1350
- /flow-inception-to-elaboration
1789
+ Ralph is useful for bounded repair loops where the success condition is objective. It is not a guarantee that the
1790
+ model will solve the task. Set wall-clock, token, tool-call, or cost limits for expensive providers.
1351
1791
 
1352
- # Iterative task execution — "iteration beats perfection"
1353
- /ralph "Fix all failing tests" --completion "npm test passes"
1792
+ ### Recursive search over large code or docs
1354
1793
 
1355
- # Long-running tasks with crash recovery (6-8 hours)
1356
- /ralph-external "Migrate to TypeScript" --completion "npx tsc --noEmit exits 0"
1794
+ ```bash
1795
+ aiwg rlm-prep docs/ --strategy semantic-boundary --size 200
1796
+ aiwg rlm-search "Where do provider quickstarts mention reload requirements?" \
1797
+ --source .aiwg/rlm-prep/<manifest-dir>/manifest.json \
1798
+ --max-parallel 4 \
1799
+ --budget 50000
1800
+ aiwg rlm-cache stats
1801
+ ```
1357
1802
 
1358
- # Process massive codebases with recursive context decomposition
1359
- /rlm-query "src/**/*.ts" "Extract all exported interfaces" --model haiku
1360
- /rlm-batch "src/components/*.tsx" "Add TypeScript types" --max-parallel 4
1803
+ Provider prompt:
1361
1804
 
1362
- # Digital forensics investigation
1363
- /forensics-investigate
1364
- /forensics-triage
1365
- /forensics-timeline
1805
+ ```text
1806
+ Find every user-facing quickstart that still tells users to manually reload after `aiwg use all`. Return file links, the quoted sentence, and the replacement language.
1807
+ ```
1366
1808
 
1367
- # Scan codebase for agent-readiness
1368
- /codebase-health --format text
1809
+ ### Dataset-backed project knowledge
1369
1810
 
1370
- # Decompose large files into agent-friendly modules
1371
- /decompose-file src/large-file.ts --execute
1811
+ ```bash
1812
+ aiwg dataset check source:docs --json
1813
+ aiwg dataset preview source:docs --count 5 --offline
1814
+ aiwg dataset query dataset:docs-index "Which provider quickstart is best for Codex?"
1815
+ ```
1372
1816
 
1373
- # Deploy to production with rollback gates
1374
- /flow-deploy-to-production
1817
+ Use dataset intelligence when the source and lineage need to be explicit. Use RLM when the immediate need is recursive
1818
+ search or fanout over files.
1375
1819
 
1376
- # Security assessment
1377
- /security-audit
1820
+ ### Session history reuse
1378
1821
 
1379
- # Voice transformation
1380
- "Apply technical-authority voice to docs/architecture.md"
1381
- "Create a voice profile based on our existing blog posts"
1822
+ ```bash
1823
+ aiwg sessions discover --workspace "$PWD" --dry-run
1824
+ aiwg sessions search "marketing audit"
1825
+ aiwg sessions extract <session-id> --format markdown
1382
1826
  ```
1383
1827
 
1384
- ---
1828
+ This is useful when prior provider conversations contain decisions that should become project artifacts. Keep import
1829
+ scope explicit and review the discovered sessions before promotion.
1385
1830
 
1386
- ## Platform Support
1831
+ ### Security, forensics, and operations prompts
1387
1832
 
1388
- AIWG supports 14 named provider integrations. Artifact support varies by provider and is adapted to each provider's native or compatibility surfaces.
1389
-
1390
- | Platform | Status | Agents | Commands | Skills | Rules | Deploy Command |
1391
- |----------|--------|--------|----------|--------|-------|---------------|
1392
- | **[Google Antigravity CLI](docs/providers/antigravity.md)** | Experimental | degraded `.agents/agents/` | — | `.agents/skills/` | `AGENTS.md` | `--provider antigravity` (alias: `agy`) |
1393
- | **Claude Code** | Tested | `.claude/agents/` | `.claude/commands/` | `.claude/skills/` | `.claude/rules/` | `aiwg use sdlc` |
1394
- | **GitHub Copilot** | Tested | `.github/agents/` | `.github/agents/` | `.github/skills/` | `.github/copilot-rules/` | `--provider copilot` |
1395
- | **Warp Terminal** | Tested | `.warp/agents/` + WARP.md | `.warp/commands/` | `.warp/skills/` | `.warp/rules/` | `--provider warp` |
1396
- | **Factory AI** | Tested | `.factory/droids/` | `.factory/commands/` | `.factory/skills/` | `.factory/rules/` | `--provider factory` |
1397
- | **Cursor** | Tested | `.cursor/agents/` | `.cursor/commands/` | `.cursor/skills/` | `.cursor/rules/` | `--provider cursor` |
1398
- | **OpenCode** | Tested | `.opencode/agent/` | `.opencode/commands/` | `.opencode/skill/` | `.opencode/rule/` | `--provider opencode` |
1399
- | **OpenAI/Codex** | Tested | `.codex/agents/` | `~/.codex/prompts/` | `.agents/skills/` | `.codex/rules/` | `--provider codex` |
1400
- | **Devin Desktop** | Tested compatibility adapter | AGENTS.md | `.windsurf/workflows/` | `.windsurf/skills/` | `.windsurf/rules/` | `--provider devin` |
1401
- | **Hermes** | Stable | — | — | `~/.hermes/skills/.aiwg/` | — | `--provider hermes` |
1402
- | **OpenClaw** | Tested | `~/.openclaw/agents/` | `~/.openclaw/commands/` | `~/.openclaw/.aiwg/skills/` | `~/.openclaw/rules/` | `--provider openclaw` |
1403
- | **OpenHuman** | Experimental | — | — | `~/.openhuman/.aiwg/skills/` | `~/.openhuman/.aiwg/rules/` | `--provider openhuman` |
1404
- | **[Pi Coding Agent](https://pi.dev/)** | Experimental | `.agents/skills/` | `.pi/prompts/` | `.agents/skills/` | `AGENTS.md` | `--provider pi` |
1405
- | **[Oh My Pi](docs/providers/omp.md)** | Experimental | `.omp/agents/` | `.omp/prompts/` | `.agents/skills/` | `.omp/AGENTS.md` | `--provider omp` |
1406
-
1407
- The legacy `--provider windsurf` selector remains supported and writes the
1408
- same `.windsurf/` compatibility paths, but new commands should use `devin`.
1409
- `devin-cli` is a distinct product surface and is not currently a deployable
1410
- AIWG provider.
1833
+ ```text
1834
+ Use the security-engineering framework to review the OAuth2 callback flow. Produce threat assumptions, concrete findings, and tests I can run.
1411
1835
 
1412
- ---
1836
+ Use the forensics framework to analyze these logs. Preserve evidence references, build a timeline, and separate confirmed facts from hypotheses.
1413
1837
 
1414
- ## CLI Reference (50 Commands)
1415
-
1416
- | Category | Commands | Description |
1417
- |----------|----------|-------------|
1418
- | **Maintenance** | `help`, `version`, `doctor`, `context-firewall`, `update` | Installation health, context safety, updates, diagnostics |
1419
- | **Framework** | `use`, `list`, `remove` | Deploy, inspect, and remove frameworks |
1420
- | **Project** | `new` | Scaffold new project with AIWG structure |
1421
- | **Workspace** | `status`, `migrate-workspace`, `rollback-workspace` | Workspace health and migration |
1422
- | **MCP** | `mcp serve`, `mcp install`, `mcp info` | Model Context Protocol server |
1423
- | **Catalog** | `catalog list`, `catalog info`, `catalog search` | Browse available extensions |
1424
- | **Marketplace packaging** | `install-plugin`, `uninstall-plugin`, `plugin-status`, `package-plugin`, `package-all-plugins` | Install and package delivery wrappers |
1425
- | **Scaffolding** | `add-agent`, `add-command`, `add-skill`, `add-template`, `scaffold-addon`, `scaffold-extension`, `scaffold-framework` | Create new extensions |
1426
- | **Ralph** | `ralph`, `ralph-status`, `ralph-abort`, `ralph-resume`, `ralph-external`, `ralph-memory`, `ralph-config` | Iterative execution engine |
1427
- | **Metrics & evidence** | `cost-report`, `cost-history`, `metrics-tokens`, `evidence` | Token usage, cost tracking, and portable evaluation evidence |
1428
- | **Index** | `index build`, `index query`, `index deps`, `index stats` | Artifact discovery and dependency graphing |
1429
- | **Documentation** | `doc-sync` | Bidirectional doc-code synchronization |
1430
- | **SDLC** | `sdlc-accelerate` | Idea-to-construction-ready pipeline |
1431
- | **Code Analysis** | `cleanup-audit` | Dead code and unused export detection |
1432
- | **Reproducibility** | `execution-mode`, `snapshot`, `checkpoint`, `reproducibility-validate` | Deterministic workflow validation |
1433
- | **Toolsmith** | `runtime-info` | Runtime environment detection |
1434
- | **Utility** | `prefill-cards`, `contribute-start`, `validate-metadata` | Development utilities |
1435
-
1436
- ### Quick Reference
1838
+ Use the ops framework to turn this production incident into a runbook update, verification checklist, and follow-up issues.
1839
+ ```
1840
+
1841
+ These prompts preserve the original README’s hands-on style while keeping provider behavior accurate: AIWG selects
1842
+ framework capabilities through discovery and provider deployment, and the deliverable remains explicit.
1843
+
1844
+ ## Platform Support
1845
+
1846
+ AIWG has registry-backed named provider integrations plus a generic fallback for tools that read Markdown context but
1847
+ do not have a dedicated adapter. The integrations share product framing: reusable project context and specialist
1848
+ workflows in the AI tools teams already use. Provider distinctions matter because each host has different native
1849
+ surfaces.
1850
+
1851
+ See [cross-platform overview](docs/integrations/cross-platform-overview.md) for the maintained comparison and setup links.
1852
+
1853
+ | Provider | Setup | Primary context | Native or conventional surfaces | Notes |
1854
+ |---|---|---|---|---|
1855
+ | Claude Code | `aiwg use all --provider claude` | `CLAUDE.md` | Skills, commands, hooks, MCP config | Best fit for rich AIWG provider packaging. See [Claude quickstart](docs/integrations/claude-code-quickstart.md). |
1856
+ | Codex | `aiwg use all --provider codex` | `AGENTS.md` | Global prompts, project skills, MCP config | Uses AGENTS.md bootstrap plus `.agents/skills/`. See [Codex quickstart](docs/integrations/codex-quickstart.md). |
1857
+ | Cursor | `aiwg use all --provider cursor` | Rules and skills | `.cursor/rules/*.mdc`, `.cursor/skills/*/SKILL.md` | Cursor rules are native; some assets remain conventional. See [Cursor quickstart](docs/integrations/cursor-quickstart.md). |
1858
+ | Windsurf | `aiwg use all --provider windsurf` | Windsurf rules | Rules and workflows | Uses Windsurf’s local rule model where available. |
1859
+ | OpenCode | `aiwg use all --provider opencode` | Agent/rule context | Provider-local agents, commands, rules | Good for lightweight terminal workflows. |
1860
+ | Gemini CLI | `aiwg use all --provider gemini` | `GEMINI.md` | Commands and context files | Keeps AIWG guidance in Gemini-readable Markdown. |
1861
+ | Qwen Code | `aiwg use all --provider qwen` | `QWEN.md` | Commands and context files | Similar Markdown-first provider packaging. |
1862
+ | Firebase Studio | `aiwg use all --provider firebase` | Studio context | Rules and generated context | Focused on Firebase Studio workspace guidance. |
1863
+ | GitHub Copilot | `aiwg use all --provider copilot` | `.github` assets | Prompts, instructions, agents, MCP config | Uses `.github/prompts/*.prompt.md`, `.github/instructions/*.instructions.md`, and `.github/agents/*.agent.md`. |
1864
+ | Devin | `aiwg use all --provider devin` | Devin-compatible context | Compatibility packaging | Uses compatibility paths where Devin can read project instructions. |
1865
+ | Factory | `aiwg use all --provider factory` | Factory context | Agents and commands where supported | Provider behavior depends on the installed Factory environment. |
1866
+ | Oh My Pi | `aiwg use all --provider omp` | `.omp/AGENTS.md` | Agents, prompts, rules, skills | Dedicated OMP quickstart: [Oh My Pi quickstart](docs/providers/omp.md). |
1867
+ | Pi Coding Agent | `aiwg use all --provider pi` | Pi context | Markdown context and provider adapters | See [Pi quickstart](docs/integrations/pi-quickstart.md). |
1868
+ | Antigravity | `aiwg use all --provider antigravity` | `AGENTS.md` and `.agents/` | Agents, skills, indexed commands, MCP config when enabled | See [Antigravity provider docs](docs/providers/antigravity.md). |
1869
+ | Generic Markdown | `aiwg use all --provider generic` | `AIWG.md` / `WORKSPACE.md` | Markdown instructions | Use when a provider reads repo docs but has no dedicated integration. |
1870
+
1871
+ After any deployment, use status and doctor before relying on the provider context:
1437
1872
 
1438
1873
  ```bash
1439
- # Deploy frameworks
1440
- aiwg use sdlc # SDLC framework
1441
- aiwg use forensics # Forensics framework
1442
- aiwg use all # Everything
1443
- aiwg use sdlc --provider copilot # Deploy to GitHub Copilot
1444
-
1445
- # Project management
1446
- aiwg new my-project # Scaffold new project
1447
- aiwg status # Workspace health
1448
- aiwg doctor # Installation diagnostics
1449
- aiwg context-firewall scan # Provider context, trust, drift, and budget audit
1450
-
1451
- # Iterative execution (Agent Loop)
1452
- aiwg ralph "Fix all tests" --completion "npm test passes"
1453
- aiwg ralph-status # Check loop progress
1454
- aiwg ralph-abort # Cancel running loop
1455
- aiwg ralph-resume # Resume interrupted loop
1456
- aiwg ralph-external "Migrate to TS" --completion "tsc --noEmit exits 0"
1457
-
1458
- # Artifact discovery
1459
- aiwg index build # Build artifact index
1460
- aiwg index query "authentication" --json
1461
- aiwg index deps .aiwg/requirements/UC-001.md --json
1462
-
1463
- # Documentation sync
1464
- aiwg doc-sync code-to-docs --dry-run
1465
- aiwg doc-sync full --interactive
1466
-
1467
- # Metrics
1468
- aiwg cost-report # Agent-native session cost breakdown
1469
- aiwg cost-report --fleet # OpenRouter per-bot MTD spend observation
1470
- aiwg evidence export --output ./evidence # Package evaluation evidence and provenance
1471
- aiwg evidence verify ./evidence # Verify hashes and the bundle checkpoint
1472
- aiwg metrics-tokens # Token usage
1473
-
1474
- # SDLC accelerate
1475
- aiwg sdlc-accelerate "Project description"
1476
- aiwg sdlc-accelerate --from-codebase .
1874
+ aiwg status --probe
1875
+ aiwg doctor
1477
1876
  ```
1478
1877
 
1479
- ---
1878
+ Fresh deployments may report that the provider should be restarted or reloaded so the host notices new files. That is
1879
+ provider-specific readiness information, not a requirement to regenerate context after every command.
1480
1880
 
1481
- ## Architecture
1881
+ ## CLI Reference
1482
1882
 
1483
- ### Extension System
1883
+ The CLI is organized around framework deployment, workspace health, discovery, governed artifacts, orchestration, and
1884
+ specialized addons. Use `aiwg help` for the current top-level surface and [CLI reference](docs/cli/reference.md) for
1885
+ the generated reference.
1484
1886
 
1485
- AIWG uses a unified extension system with 10 extension types, projected onto the supported artifact surfaces of 14 named provider integrations:
1486
-
1487
- | Type | Count | Description |
1488
- |------|-------|-------------|
1489
- | **Agents** | 188 | Specialized AI personas with defined tools, responsibilities, and operating rhythms |
1490
- | **Commands** | 50 | CLI commands and slash commands for workflow automation |
1491
- | **Skills** | 128 | Natural language workflow triggers activated by conversation patterns |
1492
- | **Rules** | 35 | Enforcement patterns deployed as consolidated index with on-demand full-rule loading |
1493
- | **Templates** | 334 | Progressive disclosure document templates for all SDLC phases |
1494
- | **Frameworks** | 8 | Complete workflow systems (SDLC, Forensics, Marketing, Research, Media Curator, Ops, Knowledge Base, Security Engineering) |
1495
- | **Addons** | 21 | Feature bundles extending frameworks (RLM, Voice, Testing Quality, UAT, Ring) |
1496
- | **Hooks** | varies | Lifecycle event handlers (pre-session, post-write, workflow tracing) |
1497
- | **Tools** | varies | External utility integrations (git, jq, npm) |
1498
- | **MCP Servers** | varies | Model Context Protocol server integrations |
1887
+ | Area | Commands | Use when |
1888
+ |---|---|---|
1889
+ | Framework deployment | `aiwg use`, `aiwg list`, `aiwg remove` | Install or remove AIWG framework/provider assets in a workspace. |
1890
+ | Getting started | `aiwg init`, `aiwg setup project`, `aiwg new`, `aiwg quickref generate` | Bootstrap a workspace or generate quick reference docs. |
1891
+ | Workspace health | `aiwg status`, `aiwg doctor`, `aiwg refresh`, `aiwg installation`, `aiwg verify` | Check readiness, repair drift, and validate generated context. |
1892
+ | Catalog and discovery | `aiwg catalog`, `aiwg discover`, `aiwg show`, `aiwg index`, `aiwg artifacts` | Find capabilities and locate generated outputs. |
1893
+ | Provider and MCP | `aiwg mcp serve`, `aiwg mcp install`, `aiwg mcp info`, `aiwg runtime-info` | Connect AIWG to MCP hosts or inspect runtime details. |
1894
+ | Execution and dispatch | `aiwg run skill`, `aiwg run script`, `aiwg output-mode`, `aiwg execution-mode` | Invoke portable skills/scripts and control output or reproducibility mode. |
1895
+ | Ralph loop | `aiwg ralph`, `aiwg ralph-status`, `aiwg ralph-resume`, `aiwg ralph-abort`, `aiwg ralph-attach` | Run bounded iterative implementation loops. |
1896
+ | Mission control | `aiwg mc start`, `aiwg mc dispatch`, `aiwg mc status`, `aiwg mc watch`, `aiwg mc stop` | Coordinate multiple bounded missions from one workspace. |
1897
+ | Sessions | `aiwg sessions discover`, `aiwg sessions import-discovered`, `aiwg sessions list`, `aiwg sessions search`, `aiwg sessions doctor` | Import, inspect, and promote provider session history. |
1898
+ | Dataset intelligence | `aiwg dataset source`, `aiwg dataset check`, `aiwg dataset preview`, `aiwg dataset plan`, `aiwg dataset ingest`, `aiwg dataset verify`, `aiwg dataset query` | Govern source intake, indexing, lineage, and queries. |
1899
+ | Evidence and metrics | `aiwg evidence export`, `aiwg evidence verify`, `aiwg cost-report --fleet` | Preserve verification records and inspect spend or usage where configured. |
1900
+ | Scaffolding | `aiwg new-bundle`, `aiwg new-extension`, `aiwg new-addon`, `aiwg new-framework`, `aiwg new-provider`, `aiwg add-agent`, `aiwg add-command`, `aiwg add-skill` | Create new AIWG packages and provider-facing assets. |
1901
+ | Issues | `aiwg issue init`, `aiwg issue plan`, `aiwg issue list`, `aiwg issue show`, `aiwg issue import`, `aiwg issue export` | Maintain local issue records and exchange snapshots with external trackers. |
1499
1902
 
1500
- ### Multi-Agent Orchestration
1903
+ Common setup and inspection flow:
1501
1904
 
1905
+ ```bash
1906
+ # Install all AIWG assets for the current provider
1907
+ aiwg use all --provider codex
1908
+
1909
+ # Verify generated context and provider readiness
1910
+ aiwg status --probe --json
1911
+ aiwg doctor
1912
+
1913
+ # Find and inspect capabilities instead of guessing command names
1914
+ aiwg discover "release planning" --limit 5
1915
+ aiwg show skill aiwg-status
1502
1916
  ```
1503
- ┌─────────────────────┐
1504
- │ Executive Orchestrator│
1505
- └──────────┬──────────┘
1506
- │
1507
- ┌────────────────┼────────────────┐
1508
- ▼ ▼ ▼
1509
- ┌──────────────┐ ┌──────────────┐ ┌──────────────┐
1510
- │Primary Author│ │ Reviewer 1 │ │ Reviewer 2 │ ← Parallel
1511
- │(e.g. Req. │ │(e.g. Security│ │(e.g. Test │
1512
- │ Analyst) │ │ Architect) │ │ Architect) │
1513
- └──────┬───────┘ └──────┬───────┘ └──────┬───────┘
1514
- └────────────────┼────────────────┘
1515
- ▼
1516
- ┌─────────────────────┐
1517
- │ Documentation │
1518
- │ Synthesizer │ ← Merge all reviews
1519
- └──────────┬──────────┘
1520
- ▼
1521
- ┌─────────────────────┐
1522
- │ Human Gate │ ← GO / NO_GO decision
1523
- └──────────┬──────────┘
1524
- ▼
1525
- ┌─────────────────────┐
1526
- │ .aiwg/ Archive │ ← Persistent artifacts
1527
- └─────────────────────┘
1917
+
1918
+ ## Architecture
1919
+
1920
+ AIWG is a portable context and workflow layer. It keeps canonical project instructions in repo-visible files, packages
1921
+ provider-specific assets for the AI tools a team uses, and preserves artifacts so work can be reviewed outside the
1922
+ original chat.
1923
+
1924
+ ```mermaid
1925
+ flowchart TD
1926
+ A[Project context<br/>WORKSPACE.md + AIWG.md] --> B[AIWG catalog]
1927
+ B --> C[Provider packaging]
1928
+ C --> D[Claude, Codex, Cursor, Copilot, Warp, OMP, Antigravity, others]
1929
+ B --> E[Specialist workflows]
1930
+ E --> F[SDLC, research, ops, security, marketing, datasets, RLM]
1931
+ E --> G[Artifacts and evidence]
1932
+ G --> H[Reports, issues, datasets, sessions, indexes]
1933
+ H --> B
1528
1934
  ```
1529
1935
 
1530
- ### YAML Metalanguage
1936
+ ### Extension System
1531
1937
 
1532
- AIWG is pioneering a declarative YAML metalanguage for multi-agent workflow orchestration. Schema-validated YAML defines agent topology, workflow DAGs, gate conditions, and artifact contracts — while natural language handles behavioral logic.
1938
+ An AIWG extension usually contains some combination of:
1533
1939
 
1534
- ```yaml
1535
- # Example: flow definition (schema-validated)
1536
- flow:
1537
- id: inception-to-elaboration
1538
- model: opus
1539
- entry_criteria:
1540
- gate: LOM
1541
- artifacts:
1542
- - path: .aiwg/requirements/vision-document.md
1543
- required: true
1544
- steps:
1545
- - id: requirements-analysis
1546
- agent: requirements-analyst
1547
- parallel_group: reviews
1548
- - id: architecture-baseline
1549
- agent: architecture-designer
1550
- parallel_group: reviews
1551
- - id: synthesis
1552
- agent: documentation-synthesizer
1553
- depends_on: [requirements-analysis, architecture-baseline]
1554
- exit_criteria:
1555
- gate: ABM
1556
- decision: [GO, CONDITIONAL_GO, NO_GO]
1557
- ```
1558
-
1559
- JSON Schema definitions for `flow.yaml`, `agent.yaml`, `rule.yaml`, and `skill.yaml` at `agentic/code/frameworks/sdlc-complete/schemas/metalanguage/`.
1560
-
1561
- ### Project Artifacts (.aiwg/)
1562
-
1563
- All SDLC artifacts persist in `.aiwg/` — structured project memory that survives across AI sessions:
1564
-
1565
- ```
1566
- .aiwg/
1567
- ├── intake/ # Project intake forms, solution profiles
1568
- ├── requirements/ # Use cases, user stories, NFRs
1569
- ├── architecture/ # SAD, ADRs, diagrams
1570
- ├── planning/ # Phase plans, iteration plans
1571
- ├── risks/ # Risk register, mitigations
1572
- ├── testing/ # Test strategy, test plans
1573
- ├── security/ # Threat models, security gates
1574
- ├── deployment/ # Deployment plans, runbooks
1575
- ├── reports/ # Generated status reports
1576
- ├── ralph/ # Agent loop state and history
1577
- └── frameworks/ # Installed framework registry
1578
- ```
1579
-
1580
- This segmentation is what makes large projects manageable. Individual code files inevitably grow, but the project knowledge stays organized into focused domains. An agent working on a deployment problem loads `@.aiwg/deployment/` and `@.aiwg/architecture/` — not the entire codebase. An agent debugging a test failure loads the relevant requirement, the test plan, and the specific source file. Context stays sharp regardless of project size.
1581
-
1582
- `aiwg index` amplifies this further — it builds a searchable artifact index so agents resolve lookups in a single query instead of browsing. Without tooling: 3-6 documents to find what's needed. With AIWG structure: 2-3. With the index: usually 1.
1940
+ | Asset | Role |
1941
+ |---|---|
1942
+ | Agents | Persistent role definitions, responsibilities, and routing constraints. |
1943
+ | Skills | Task-specific procedures with triggers, inputs, outputs, and evidence rules. |
1944
+ | Commands | Provider-facing shortcuts or prompt templates. |
1945
+ | Rules | Policies and reusable constraints. |
1946
+ | Schemas | Structured contracts for plans, artifacts, manifests, and reports. |
1947
+ | Templates | Repeatable starting points for generated files. |
1948
+ | Scripts | Local deterministic helpers used by workflows. |
1583
1949
 
1584
- ---
1950
+ The registry and discovery index make those assets findable without requiring every provider to support every asset
1951
+ type natively. When a provider lacks a native concept, AIWG packages the asset as Markdown context or a conventional
1952
+ file the provider can read.
1585
1953
 
1586
- ## Agent Loop — Autonomous Long-Running Agent Orchestration
1954
+ ### Multi-Agent Orchestration
1587
1955
 
1588
- The Agent Loop is the core execution philosophy: **iteration beats perfection**. Instead of getting everything right on the first attempt, the agent executes in a retry loop where errors become learning data. Ralph supports both in-session loops and **crash-resilient external loops that run indefinitely** — surviving process crashes, terminal disconnects, and system reboots.
1956
+ AIWG’s orchestration model is explicit about roles and handoffs. A complex task can move through a steward, specialist
1957
+ skill, review step, evidence export, and follow-up issue without losing the artifact trail.
1589
1958
 
1590
- ### In-Session Ralph (Minutes to Hours)
1959
+ ```mermaid
1960
+ sequenceDiagram
1961
+ participant U as User
1962
+ participant S as Steward / Discover
1963
+ participant W as Specialist Workflow
1964
+ participant P as Provider Tooling
1965
+ participant E as Evidence Store
1591
1966
 
1592
- ```bash
1593
- # Iterative task execution with automatic error recovery
1594
- /ralph "Fix all failing tests" --completion "npm test passes with 0 failures"
1595
- /ralph "Reach 80% coverage" --completion "coverage report shows >80%" --max-iterations 20
1967
+ U->>S: Describe outcome and success check
1968
+ S->>W: Select capability and inputs
1969
+ W->>P: Execute provider-local or CLI steps
1970
+ P-->>W: Results, files, diagnostics
1971
+ W->>E: Write report/evidence/issue refs
1972
+ W-->>U: Deliverable with checks and next step
1973
+ ```
1974
+
1975
+ This structure is why AIWG documentation emphasizes “first useful task, concrete deliverable, success check, and next
1976
+ step.” It gives a model enough direction to act while leaving a reviewer enough evidence to verify the result.
1977
+
1978
+ ### YAML Metalanguage
1979
+
1980
+ Many AIWG assets use YAML frontmatter or YAML schemas so capabilities can be discovered, validated, and converted
1981
+ between provider formats.
1596
1982
 
1597
- # Issue-driven Ralph — posts cycle status to issue threads, incorporates human feedback
1598
- /issue-driven-ralph 42 # Drives issue #42 with 2-way human-AI collaboration
1983
+ ```yaml
1984
+ ---
1985
+ namespace: aiwg
1986
+ name: release-notes
1987
+ description: Draft release notes from approved issue and changelog artifacts
1988
+ platforms: [all]
1989
+ triggers:
1990
+ - release notes
1991
+ - changelog summary
1992
+ outputs:
1993
+ - docs/releases/{version}.md
1994
+ evidence:
1995
+ required:
1996
+ - source_issue_refs
1997
+ - changelog_refs
1998
+ ---
1599
1999
  ```
1600
2000
 
1601
- ### External Ralph — Crash-Resilient Autonomous Agents (Hours to Days)
2001
+ The metadata is not decorative. It lets `aiwg discover` find the capability, lets validation detect missing fields,
2002
+ and gives provider adapters enough information to package the asset accurately.
1602
2003
 
1603
- External Ralph runs as a **persistent background process** with PID file tracking, crash recovery, and automatic restart. The agent continues working even if your terminal disconnects or the host reboots.
2004
+ ### Project Artifacts
1604
2005
 
1605
- ```bash
1606
- # Long-running autonomous task (6-8+ hours, survives crashes)
1607
- /ralph-external "Migrate entire codebase to TypeScript" \
1608
- --completion "npx tsc --noEmit exits 0" \
1609
- --timeout 480
2006
+ AIWG-generated artifacts are intentionally ordinary files: Markdown, JSON, YAML, SQLite-backed local stores when
2007
+ enabled, and provider-readable context. That makes them inspectable in a code review and portable across machines.
1610
2008
 
1611
- # Autonomous code review loop
1612
- /ralph-external "Review and fix all security vulnerabilities" \
1613
- --completion "npm audit shows 0 vulnerabilities"
2009
+ Examples include:
1614
2010
 
1615
- # Continuous integration loop
1616
- /ralph-external "Get all tests passing on Node 18 and 22" \
1617
- --completion "npm test passes on both versions"
2011
+ ```text
2012
+ .aiwg/reports/context-firewall-*.md
2013
+ .aiwg/reports/doc-sync-audit-*.md
2014
+ .aiwg/evidence/*.json
2015
+ .aiwg/issues/*.json
2016
+ .aiwg/datasets/**/manifest.json
2017
+ .aiwg/rlm-prep/**/manifest.json
2018
+ .aiwg/sessions/**
1618
2019
  ```
1619
2020
 
1620
- External Ralph features:
2021
+ Artifact indexes reduce manual browsing, but they do not replace source review. Treat indexed results as navigation
2022
+ aids with links back to the original files.
1621
2023
 
1622
- - **Crash resilience** — PID file recovery, automatic restart on process death
1623
- - **Checkpoint system** — saves progress at each iteration boundary, resumes from last checkpoint
1624
- - **Cross-session persistence** — state stored in `.aiwg/ralph-external/`, survives terminal disconnects
1625
- - **Debug memory** — learns from failure patterns across iterations, applies lessons to subsequent attempts
1626
- - **Episodic memory** — `/ralph-reflect` shows accumulated learnings and strategy evolution
1627
- - **Completion reports** — detailed iteration history saved to `.aiwg/ralph/`
2024
+ ## Agent Loop — Autonomous Long-Running Agent Orchestration
1628
2025
 
1629
- ### Scheduled and Remote Agents
2026
+ Ralph is AIWG’s bounded iterative agent loop. It is intended for tasks where the objective and completion criterion
2027
+ can be checked: fixing tests, applying a migration, updating docs to match a report, or carrying a refactor through
2028
+ verification.
1630
2029
 
1631
2030
  ```bash
1632
- # Schedule recurring autonomous agent tasks
1633
- /schedule create "Run security audit" --cron "0 9 * * 1" # Every Monday 9am
1634
- /schedule create "Check dependency updates" --cron "0 0 * *" # Monthly
2031
+ aiwg ralph "Update provider quickstarts from the marketing audit" \
2032
+ --completion "changed files match the approved audit scope and markdown links pass" \
2033
+ --max-iterations 6 \
2034
+ --max-wall-clock-minutes 60 \
2035
+ --max-tool-calls 120
2036
+ ```
1635
2037
 
1636
- # Remote agent triggers — execute on schedule from anywhere
1637
- /schedule list
1638
- /schedule run <trigger-id>
2038
+ Ralph records loop state so work can be inspected and resumed when supported by the selected provider and local environment.
2039
+
2040
+ ```bash
2041
+ aiwg ralph-status
2042
+ aiwg ralph-attach <loop-id>
2043
+ aiwg ralph-resume <loop-id>
2044
+ aiwg ralph-abort <loop-id>
1639
2045
  ```
1640
2046
 
1641
- ### Ralph Control
2047
+ Use budgets for any loop that may call a remote model or external provider:
1642
2048
 
1643
2049
  ```bash
1644
- /ralph-status # Check current/previous loop status
1645
- /ralph-resume # Resume interrupted loop from last checkpoint
1646
- /ralph-abort # Cancel running loop (optionally revert changes)
1647
- /ralph-memory # View debug memory entries and failure patterns
1648
- /ralph-reflect # View episodic memory and strategy evolution
1649
- /ralph-analytics # Execution metrics and performance history
2050
+ aiwg ralph "Reduce flaky integration tests" \
2051
+ --completion "the flaky-test reproduction passes 10 consecutive runs" \
2052
+ --max-total-tokens 200000 \
2053
+ --max-total-cost 10 \
2054
+ --budget-stop-policy budget-wins
1650
2055
  ```
1651
2056
 
1652
- ### How It Works
2057
+ Long-running automation should still produce reviewable outputs: changed files, reports, evidence, status logs, and
2058
+ the exact checks run. Ralph can continue work within configured limits, but it cannot guarantee a solution, fixed
2059
+ runtime, or provider availability.
1653
2060
 
1654
- Each iteration follows the TAO loop (Thought → Action → Observation):
2061
+ Mission Control builds on the same principle for multiple bounded work items:
1655
2062
 
1656
- ```
1657
- Iteration N:
1658
- 1. THINK — Analyze current state + accumulated learnings from iterations 1..N-1
1659
- 2. ACT — Make changes based on task + debug memory + failure patterns
1660
- 3. VERIFY — Run completion command (tests, build, lint, coverage, etc.)
1661
- 4. LEARN — If verification fails, extract root cause → store in debug memory
1662
- 5. DECIDE — Pass? → Complete. Fail? → Iterate. Max retries? → Escalate to human.
2063
+ ```bash
2064
+ aiwg mc start --name "docs audit follow-up" --max-missions 4
2065
+ aiwg mc dispatch <session-id> \
2066
+ "Validate README command examples" \
2067
+ --completion "all documented commands are current or labeled provider-specific"
2068
+ aiwg mc run <session-id>
2069
+ aiwg mc status <session-id>
2070
+ aiwg mc watch <session-id>
1663
2071
  ```
1664
2072
 
1665
- The debug memory system implements executable feedback: the agent doesn't just retry — it learns *what went wrong* and *why*, then applies that knowledge to the next attempt. After 3 failed attempts at the same root cause, it escalates to a human rather than looping forever.
2073
+ ## RLM — Recursive Context Decomposition
1666
2074
 
1667
- Research foundation: Self-Refine (Madaan et al., NeurIPS 2023), ReAct (Yao et al., ICLR 2023), METR 2025 (recovery capability dominates agentic task success), Reflexion (Shinn et al., 2023).
2075
+ For a bounded batch, specify the parallelism limit explicitly:
1668
2076
 
1669
- ---
2077
+ ```text
2078
+ /rlm-batch "src/components/*.tsx" "Add TypeScript types" --max-parallel 4
2079
+ ```
1670
2080
 
1671
- ## RLM — Recursive Context Decomposition
2081
+ RLM helps with sources that are too large to fit comfortably in one model context. It prepares files into traceable
2082
+ chunks, fans a query out across those chunks, and merges results with links back to the source material.
2083
+
2084
+ ```mermaid
2085
+ flowchart LR
2086
+ A[Source files] --> B[rlm-prep]
2087
+ B --> C[manifest.json]
2088
+ C --> D[rlm-search / fanout]
2089
+ D --> E[ranked findings]
2090
+ E --> F[source-linked answer]
2091
+ C --> G[rlm-cache]
2092
+ ```
1672
2093
 
1673
- Process codebases and documents far beyond any model's context window:
2094
+ Use `rlm-prep` when you want to prepare a file tree once and reuse it for several searches.
1674
2095
 
1675
2096
  ```bash
1676
- # Query: fan-out across files, gather results
1677
- /rlm-query "src/**/*.ts" "Extract all exported interfaces" --model haiku
2097
+ # Prepare source or docs for recursive search
2098
+ aiwg rlm-prep src/ --strategy semantic-boundary --size 200 --overlap 20
2099
+ aiwg rlm-prep docs/ --strategy fixed-count --size 150
1678
2100
 
1679
- # Batch: parallel processing with configurable concurrency
1680
- /rlm-batch "src/components/*.tsx" "Add TypeScript types" --max-parallel 4
2101
+ # Search the prepared source
2102
+ aiwg rlm-search "Where is provider reload status calculated?" \
2103
+ --source .aiwg/rlm-prep/<source-hash>/manifest.json \
2104
+ --depth 3 \
2105
+ --max-parallel 4 \
2106
+ --budget 50000
2107
+
2108
+ # Run a direct fanout query over a manifest or chunks directory
2109
+ aiwg fanout "Summarize every stale quickstart command" \
2110
+ --chunks .aiwg/rlm-prep/<source-hash>/manifest.json \
2111
+ --parallel 4
1681
2112
 
1682
- # Status: monitor decomposition progress
1683
- /rlm-status
2113
+ # Inspect cache state
2114
+ aiwg rlm-status
2115
+ aiwg rlm-cache stats
1684
2116
  ```
1685
2117
 
1686
- The RLM addon decomposes large inputs into chunks, delegates each to a sub-agent, and synthesizes results. Processes 10M+ tokens through recursive delegation.
2118
+ Use `chunk` for a single-file manual workflow:
1687
2119
 
1688
- Research foundation: Recursive Language Models (Zhang, Kraska, Khattab — MIT CSAIL, 2026).
2120
+ ```bash
2121
+ aiwg chunk README.md --size 200 --overlap 20 --format json --output .aiwg/chunks/readme
2122
+ ```
1689
2123
 
1690
- ---
2124
+ RLM is a retrieval and decomposition workflow, not a magic context override. Quality depends on chunk boundaries,
2125
+ source coverage, prompt specificity, model capability, and budget. For high-stakes review, ask for quoted source
2126
+ links, inspect the cited chunks, and rerun targeted searches for disputed claims.
1691
2127
 
1692
2128
  ## Research Foundations
1693
2129
 
1694
- AIWG's architecture is grounded in peer-reviewed research across cognitive science, multi-agent systems, software engineering, and AI safety. Reference summaries live in `docs/references/` (REF-NNN entries), ordered highest to lowest GRADE evidence quality within each category.
2130
+ AIWG draws design ideas from research across cognitive science, multi-agent
2131
+ systems, software engineering, retrieval, provenance, and AI safety. The
2132
+ results cited below belong to the referenced papers or systems; they are not
2133
+ AIWG performance guarantees. Reference summaries live in `docs/references/`
2134
+ (REF-NNN entries). The bibliography groups related design topics.
1695
2135
 
1696
2136
  ### Cognitive Foundations
1697
2137
 
@@ -1705,22 +2145,31 @@ AIWG's architecture is grounded in peer-reviewed research across cognitive scien
1705
2145
  ### Multi-Agent Systems & Orchestration
1706
2146
 
1707
2147
  - Jacobs, R.A. et al. (1991). [Adaptive Mixtures of Local Experts](https://doi.org/10.1162/neco.1991.3.1.79). *Neural Computation*, 3(1), 79–87. (Mixture-of-Experts foundation)
1708
- - Hong, S. et al. (2024). [MetaGPT: Meta Programming for a Multi-Agent Collaborative Framework](https://arxiv.org/abs/2308.00352). *ICLR 2024*. (85.9% HumanEval, SOP-based orchestration)
2148
+ - Hong, S. et al. (2024). [MetaGPT: Meta Programming for a Multi-Agent Collaborative
2149
+ Framework](https://arxiv.org/abs/2308.00352). *ICLR 2024*.
1709
2150
  - Qian, C. et al. (2024). [ChatDev: Communicative Agents for Software Development](https://arxiv.org/abs/2307.07924). *ACL 2024*.
1710
2151
  - Shen, Y. et al. (2023). [HuggingGPT: Solving AI Tasks with ChatGPT and its Friends in HuggingFace](https://arxiv.org/abs/2303.17580). *NeurIPS 2023*.
1711
2152
  - Tao, W. et al. (2024). [MAGIS: LLM-Based Multi-Agent Framework for GitHub Issue Resolution](https://arxiv.org/abs/2403.17927).
1712
- - Zhang, J. et al. (2025). [AFlow: Automating Agentic Workflow Generation](https://arxiv.org/abs/2410.10762). *ICLR 2025 Oral*. (5.7% avg gain over best manual methods)
2153
+ - Zhang, J. et al. (2025). [AFlow: Automating Agentic Workflow Generation](https://arxiv.org/abs/2410.10762). *ICLR
2154
+ 2025 Oral*.
1713
2155
  - Wu, Q. et al. (2023). [AutoGen: Enabling Next-Gen LLM Applications via Multi-Agent Conversation](https://arxiv.org/abs/2308.08155). (Conversational multi-agent framework)
1714
2156
  - Yu, C. et al. (2025). [A Survey on Agent Workflow — Status and Future](https://arxiv.org/abs/2508.01186). (24 systems, 11 metrics)
1715
- - Lodha, D. et al. (2026). [MCP-Diag: A Deterministic, Protocol-Driven Architecture for AI-Native Network Diagnostics](https://arxiv.org/abs/2601.22633). *COMSNETS 2026*. (First production MCP system)
1716
- - Yu, G. (2026). [AdaptOrch: Adaptive Orchestration for Multi-Agent LLM Systems Through Topology-Aware Task Planning](https://arxiv.org/abs/2502.09340). (12–23% improvement across 4 topologies)
2157
+ - Lodha, D. et al. (2026). [MCP-Diag: A Deterministic, Protocol-Driven Architecture for AI-Native Network
2158
+ Diagnostics](https://arxiv.org/abs/2601.22633). *COMSNETS 2026*.
2159
+ - Yu, G. (2026). [AdaptOrch: Adaptive Orchestration for Multi-Agent LLM Systems Through Topology-Aware Task Planning](https://arxiv.org/abs/2502.09340).
1717
2160
  - Gerred (2025). [Multi-Agent Orchestration](https://gerred.github.io/building-an-agentic-system/second-edition/part-iv-advanced-patterns/chapter-10-multi-agent-orchestration.html). Tool isolation, resource boundaries, observable coordination.
1718
2161
  - Falconer, S. (2025). [Event-Driven Multi-Agent Systems](https://www.confluent.io/blog/event-driven-multi-agent-systems/). Confluent. 4 Kafka orchestration patterns.
1719
2162
  - Mario, M. (2025). [Multi-Agent System Patterns: A Unified Guide to Designing Agentic Architectures](https://medium.com/@mjgmario/multi-agent-system-patterns-a-unified-guide-to-designing-agentic-architectures-04bb31ab9c41). 4-dimensional framework.
1720
- - Runkle, S. (2026). [Choosing the Right Multi-Agent Architecture](https://www.blog.langchain.com/choosing-the-right-multi-agent-architecture/). LangChain. Subagents, skills, handoffs, 90.2% improvement stat.
1721
- - Towards Data Science (2025). [Why Your Multi-Agent System Is Failing: Escaping the 17x Error Trap](https://towardsdatascience.com/why-your-multi-agent-system-is-failing-escaping-the-17x-error-trap-of-the-bag-of-agents/). 17.2x error amplification, 4-agent coordination threshold.
2163
+ - Runkle, S. (2026). [Choosing the Right Multi-Agent
2164
+ Architecture](https://www.blog.langchain.com/choosing-the-right-multi-agent-architecture/). LangChain. Subagents,
2165
+ skills, and handoffs.
2166
+ - Towards Data Science (2025). [Why Your Multi-Agent System Is Failing: Escaping the 17x Error
2167
+ Trap](https://towardsdatascience.com/why-your-multi-agent-system-is-failing-escaping-the-17x-error-trap-of-the-bag-of-agents/).
2168
+ Coordination failure analysis.
1722
2169
  - NexAI Tech (2025). [Multi-AI Agent Architecture Patterns for Scale](https://nexaitech.com/multi-ai-agent-architecutre-patterns-for-scale/). Enterprise 5-layer architecture, 3 orchestration patterns.
1723
- - Wexford, E. (2026). [How to Build Multi-Agent Systems: Complete 2026 Guide](https://dev.to/eira-wexford/how-to-build-multi-agent-systems-complete-2026-guide-1io6). DEV Community. 3–7 agents optimal sizing.
2170
+ - Wexford, E. (2026). [How to Build Multi-Agent Systems: Complete 2026
2171
+ Guide](https://dev.to/eira-wexford/how-to-build-multi-agent-systems-complete-2026-guide-1io6). DEV Community.
2172
+ Multi-agent design guidance.
1724
2173
 
1725
2174
  ### Reasoning & Planning
1726
2175
 
@@ -1730,11 +2179,13 @@ AIWG's architecture is grounded in peer-reviewed research across cognitive scien
1730
2179
  - Yao, S. et al. (2023). [Tree of Thoughts: Deliberate Problem Solving with Large Language Models](https://arxiv.org/abs/2305.10601). *NeurIPS 2023*.
1731
2180
  - Zhou, A. et al. (2024). [Language Agent Tree Search Unifies Reasoning, Acting, and Planning in Language Models](https://arxiv.org/abs/2310.04406). *ICML 2024*.
1732
2181
  - Kojima, T. et al. (2022). [Large Language Models are Zero-Shot Reasoners](https://arxiv.org/abs/2205.11916). *NeurIPS 2022*. ("Let's think step by step")
1733
- - Liu, Z. et al. (2026). [Exploratory Memory-Augmented LLM Agent via Hybrid On- and Off-Policy Optimization (EMPO²)](https://arxiv.org/abs/2602.23008). *ICLR 2026*. (128.6% over GRPO on ScienceWorld)
2182
+ - Liu, Z. et al. (2026). [Exploratory Memory-Augmented LLM Agent via Hybrid On- and Off-Policy Optimization
2183
+ (EMPO²)](https://arxiv.org/abs/2602.23008). *ICLR 2026*.
1734
2184
 
1735
2185
  ### Self-Correction & Iterative Refinement
1736
2186
 
1737
- - Madaan, A. et al. (2023). [Self-Refine: Iterative Refinement with Self-Feedback](https://arxiv.org/abs/2303.17651). *NeurIPS 2023*. (+4.2% HumanEval, −63% revision cost)
2187
+ - Madaan, A. et al. (2023). [Self-Refine: Iterative Refinement with Self-Feedback](https://arxiv.org/abs/2303.17651).
2188
+ *NeurIPS 2023*.
1738
2189
  - Shinn, N. et al. (2023). [Reflexion: Language Agents with Verbal Reinforcement Learning](https://arxiv.org/abs/2303.11366). *NeurIPS 2023*.
1739
2190
 
1740
2191
  ### Stage-Gate, SDLC & Traceability
@@ -1746,45 +2197,49 @@ AIWG's architecture is grounded in peer-reviewed research across cognitive scien
1746
2197
  ### Software Engineering & Agent-Computer Interface
1747
2198
 
1748
2199
  - Jimenez, C.E. et al. (2024). [SWE-bench: Can Language Models Resolve Real-world GitHub Issues?](https://www.swebench.com). *ICLR 2024*.
1749
- - Wang, X. et al. (2024). [Executable Code Actions Elicit Better LLM Agents (CodeAct)](https://arxiv.org/abs/2402.01030). *ICML 2024*. (Up to 20% higher success rate)
1750
- - Yang, J. et al. (2024). [SWE-agent: Agent-Computer Interfaces Enable Automated Software Engineering](https://arxiv.org/abs/2405.15793). *NeurIPS 2024*. (12.47% SWE-bench)
1751
- - Laurent, A. (2025). [A Comparison of AI Code Assistants for Large Codebases](https://intuitionlabs.ai/articles/ai-code-assistants-large-codebases). IntuitionLabs. (62% AI code contains flaws)
1752
- - Augment Code (2025). [AI Coding Assistants for Large Codebases: A Complete Guide](https://www.augmentcode.com/tools/ai-coding-assistants-for-large-codebases-a-complete-guide). (73% compile locally but violate patterns)
2200
+ - Wang, X. et al. (2024). [Executable Code Actions Elicit Better LLM Agents
2201
+ (CodeAct)](https://arxiv.org/abs/2402.01030). *ICML 2024*.
2202
+ - Yang, J. et al. (2024). [SWE-agent: Agent-Computer Interfaces Enable Automated Software
2203
+ Engineering](https://arxiv.org/abs/2405.15793). *NeurIPS 2024*.
2204
+ - Laurent, A. (2025). [A Comparison of AI Code Assistants for Large
2205
+ Codebases](https://intuitionlabs.ai/articles/ai-code-assistants-large-codebases). IntuitionLabs.
2206
+ - Augment Code (2025). [AI Coding Assistants for Large Codebases: A Complete Guide](https://www.augmentcode.com/tools/ai-coding-assistants-for-large-codebases-a-complete-guide).
1753
2207
  - AlgoMaster (2025). [How to Use AI Effectively in Large Codebases](https://blog.algomaster.io/p/using-ai-effectively-in-large-codebases). Retrieval as bottleneck framing.
1754
2208
 
1755
2209
  ### Context Engineering & Memory
1756
2210
 
1757
2211
  - Liu, N.F. et al. (2024). [Lost in the Middle: How Language Models Use Long Contexts](https://arxiv.org/abs/2307.03172). *TACL* 12, 157–173. doi:10.1162/tacl_a_00638
1758
2212
  - Dai, Y. et al. (2025). [Pretraining Context Compressor for Large Language Models with Embedding-Based Memory](https://aclanthology.org/2025.acl-long.1394.pdf). *ACL 2025*.
1759
- - Kang, M. et al. (2025). [ACON: Optimizing Context Compression for Long-Horizon LLM Agents](https://arxiv.org/abs/2510.00615). (26–54% peak token reduction, >95% accuracy preserved)
1760
- - Liu, F. & Qiu, H. (2025). [Context Cascade Compression (C3): Exploring the Upper Limits of Text Compression](https://arxiv.org/abs/2511.15244). (98% precision at 20x compression)
2213
+ - Kang, M. et al. (2025). [ACON: Optimizing Context Compression for Long-Horizon LLM Agents](https://arxiv.org/abs/2510.00615).
2214
+ - Liu, F. & Qiu, H. (2025). [Context Cascade Compression (C3): Exploring the Upper Limits of Text Compression](https://arxiv.org/abs/2511.15244).
1761
2215
  - Vasilopoulos, A. (2026). [Codified Context: Infrastructure for AI Agents in a Complex Codebase](https://arxiv.org/abs/2602.20478). (Three-tier context infrastructure: constitution + 19 agents + 34-doc KB)
1762
2216
  - Ostby, D.L. (2025). [Stingy Context: Compressing Code Context for Cost-Effective AI Development Assistance](https://arxiv.org/abs/2512.15504). (TREEFRAG, 18:1 compression ratio)
1763
2217
  - Anthropic Applied AI Team (2026). [Effective Context Engineering for AI Agents](https://www.anthropic.com/engineering/effective-context-engineering-for-ai-agents). Anthropic Engineering Blog.
1764
- - Huang, J.Y. et al. (2026). [Do LLMs Benefit From Their Own Words?](https://arxiv.org/abs/2602.24287) (36.4% of multi-turn prompts self-contained; up to 10x context reduction)
2218
+ - Huang, J.Y. et al. (2026). [Do LLMs Benefit From Their Own Words?](https://arxiv.org/abs/2602.24287)
1765
2219
  - Böckeler, B. (2026). [Context Engineering for Coding Agents](https://martinfowler.com/articles/exploring-gen-ai/context-engineering-coding-agents.html). Martin Fowler's Blog. Two-category framework.
1766
- - Haseeb, M. (2025). [Context Engineering for Multi-Agent LLM Code Assistants](https://arxiv.org/abs/2508.08322). (80% vs 40% single-shot success)
1767
- - Verma, N. (2026). [Focus Agent: LLM Agent with Active Context Compression for SWE-Bench](https://arxiv.org/abs/2501.09067). (22.7% token reduction via consolidate/withdraw)
2220
+ - Haseeb, M. (2025). [Context Engineering for Multi-Agent LLM Code Assistants](https://arxiv.org/abs/2508.08322).
2221
+ - Verma, N. (2026). [Focus Agent: LLM Agent with Active Context Compression for SWE-Bench](https://arxiv.org/abs/2501.09067).
1768
2222
  - Zylos Research (2026). [Long-Running AI Agents and Task Decomposition](https://zylos.ai/research/2026-01-16-long-running-ai-agents). (35-min degradation threshold, Planner-Worker model)
1769
- - Zylos Research (2026). [LLM Context Window Management and Long-Context Strategies](https://zylos.ai/research/2026-01-19-llm-context-management). (Lost-in-Middle persists, TTT-E2E 35× speedup)
2223
+ - Zylos Research (2026). [LLM Context Window Management and Long-Context Strategies](https://zylos.ai/research/2026-01-19-llm-context-management).
1770
2224
 
1771
2225
  ### Agent Memory & Knowledge Systems
1772
2226
 
1773
2227
  - Laird, J.E. et al. (1987). [SOAR: An Architecture for General Intelligence](https://doi.org/10.1016/0004-3702(87)90050-6). *Artificial Intelligence*, 33(1), 1–64.
1774
2228
  - Anderson, J.R. et al. (2004). [An Integrated Theory of the Mind (ACT-R)](https://doi.org/10.1037/0033-295X.111.4.1036). *Psychological Review*, 111(4), 1036–1060.
1775
2229
  - Park, J.S. et al. (2023). [Generative Agents: Interactive Simulacra of Human Behavior](https://arxiv.org/abs/2304.03442). *UIST 2023*. doi:10.1145/3586183.3606763
1776
- - Xu, W. et al. (2025). [A-MEM: Agentic Memory for LLM Agents](https://arxiv.org/abs/2502.12110). (Zettelkasten-inspired, 85–93% token reduction)
2230
+ - Xu, W. et al. (2025). [A-MEM: Agentic Memory for LLM Agents](https://arxiv.org/abs/2502.12110).
1777
2231
  - Hu, Y. et al. (2025). [Memory in the Age of AI Agents: A Survey](https://arxiv.org/abs/2512.13564). (Surveys 100+ implementations, forms-functions-dynamics framework)
1778
- - Rezazadeh, A. et al. (2025). [Collaborative Memory: Multi-User Memory Sharing in LLM Agents with Dynamic Access Control](https://arxiv.org/abs/2505.18279). (61% resource reduction)
2232
+ - Rezazadeh, A. et al. (2025). [Collaborative Memory: Multi-User Memory Sharing in LLM Agents with Dynamic Access Control](https://arxiv.org/abs/2505.18279).
1779
2233
  - Yuen, S. et al. (2025). [Intrinsic Memory Agents: Heterogeneous Multi-Agent LLM Systems through Structured Contextual Memory](https://arxiv.org/abs/2508.08997). Role-aligned heterogeneous memory.
1780
2234
  - Graves, A., Wayne, G. & Danihelka, I. (2014). [Neural Turing Machines](https://arxiv.org/abs/1410.5401). External memory architectures.
1781
2235
  - Packer, C. et al. (2023). [MemGPT: Towards LLMs as Operating Systems](https://arxiv.org/abs/2310.08560). OS-inspired virtual context paging.
1782
2236
  - Yu, Z. et al. (2026). [Multi-Agent Memory from a Computer Architecture Perspective](https://arxiv.org/abs/2603.10062). *Architecture 2.0 '26*. Three-layer I/O-cache-memory hierarchy.
1783
- - Chhikara, P. et al. (2025). [Mem0: Building Production-Ready AI Agents with Scalable Long-Term Memory](https://arxiv.org/abs/2504.19413). (26% accuracy gain, 91% latency reduction)
2237
+ - Chhikara, P. et al. (2025). [Mem0: Building Production-Ready AI Agents with Scalable Long-Term Memory](https://arxiv.org/abs/2504.19413).
1784
2238
 
1785
2239
  ### Recursive Context Decomposition
1786
2240
 
1787
- - Zhang, A.L., Kraska, T. & Khattab, O. (2026). [Recursive Language Models](https://arxiv.org/abs/2512.24601). *arXiv:2512.24601*. MIT CSAIL. (10M+ token processing, up to 3x cheaper than summarization)
2241
+ - Zhang, A.L., Kraska, T. & Khattab, O. (2026). [Recursive Language Models](https://arxiv.org/abs/2512.24601).
2242
+ *arXiv:2512.24601*. MIT CSAIL.
1788
2243
 
1789
2244
  ### Provenance, Reproducibility & Research Management
1790
2245
 
@@ -1792,8 +2247,8 @@ AIWG's architecture is grounded in peer-reviewed research across cognitive scien
1792
2247
  - W3C (2013). [PROV-DM: The PROV Data Model](https://www.w3.org/TR/prov-dm/). W3C Recommendation.
1793
2248
  - CCSDS (2024). [Reference Model for an Open Archival Information System (OAIS)](https://public.ccsds.org/Pubs/650x0m2.pdf). ISO 14721. (Digital preservation lifecycle)
1794
2249
  - GRADE Working Group (2004–present). [GRADE Handbook](https://www.gradeworkinggroup.org/). Evidence quality assessment. Adopted by WHO, Cochrane, NICE, and 100+ organizations.
1795
- - Schmidgall, S. et al. (2025). [Agent Laboratory: Using LLM Agents as Research Assistants](https://arxiv.org/abs/2501.04227). (84% cost reduction)
1796
- - Sureshkumar, V. et al. (2026). [R-LAM: Towards Reproducibility in Large Action Model Workflows](https://arxiv.org/abs/2601.09749). (47% of workflows non-reproducible without constraints)
2250
+ - Schmidgall, S. et al. (2025). [Agent Laboratory: Using LLM Agents as Research Assistants](https://arxiv.org/abs/2501.04227).
2251
+ - Sureshkumar, V. et al. (2026). [R-LAM: Towards Reproducibility in Large Action Model Workflows](https://arxiv.org/abs/2601.09749).
1797
2252
  - ServiceNow Research (2025). LitLLM for Scientific Literature Reviews. RAG-based literature review, no hallucination approach.
1798
2253
 
1799
2254
  ### AI Safety & Failure Modes
@@ -1833,81 +2288,111 @@ AIWG's architecture is grounded in peer-reviewed research across cognitive scien
1833
2288
 
1834
2289
  ### Constrained Generation & Output Validation
1835
2290
 
1836
- - Beurer-Kellner, L., Fischer, M. & Vechev, M. (2023). [Prompting Is Programming: A Query Language for Large Language Models (LMQL)](https://arxiv.org/abs/2212.06094). *PLDI 2023*. doi:10.1145/3591300 (26–85% token reduction)
1837
- - Willard, B.T. & Louf, R. (2023). [Efficient Guided Generation for Large Language Models (Outlines)](https://arxiv.org/abs/2307.09702). (0% parse failures by construction)
1838
- - Lhoest, Q. & Turuta, M. (2024). [Structured Generation with Outlines](https://huggingface.co/blog/outlines-structured-generation). Hugging Face Blog. (1.5–3x speedup)
2291
+ - Beurer-Kellner, L., Fischer, M. & Vechev, M. (2023). [Prompting Is Programming: A Query Language for Large Language
2292
+ Models (LMQL)](https://arxiv.org/abs/2212.06094). *PLDI 2023*. doi:10.1145/3591300
2293
+ - Willard, B.T. & Louf, R. (2023). [Efficient Guided Generation for Large Language Models (Outlines)](https://arxiv.org/abs/2307.09702).
2294
+ - Lhoest, Q. & Turuta, M. (2024). [Structured Generation with
2295
+ Outlines](https://huggingface.co/blog/outlines-structured-generation). Hugging Face Blog.
1839
2296
  - Gerganov, G. et al. (2024). [Grammar-Based Sampling (GBNF) — llama.cpp](https://github.com/ggerganov/llama.cpp/blob/master/grammars/README.md). Context-free grammar constrained sampling.
1840
2297
 
1841
2298
  ### LLM Serving & Local Deployment
1842
2299
 
1843
- - Yu, G. et al. (2022). [Orca: A Distributed Serving System for Transformer-Based Generative Models](https://www.usenix.org/conference/osdi22/presentation/yu). *OSDI '22*. (36.9x throughput improvement)
1844
- - Kwon, W. et al. (2023). [Efficient Memory Management for Large Language Model Serving with PagedAttention](https://arxiv.org/abs/2309.06180). *SOSP '23*. UC Berkeley. (2–4x throughput vs HuggingFace)
2300
+ - Yu, G. et al. (2022). [Orca: A Distributed Serving System for Transformer-Based Generative
2301
+ Models](https://www.usenix.org/conference/osdi22/presentation/yu). *OSDI '22*.
2302
+ - Kwon, W. et al. (2023). [Efficient Memory Management for Large Language Model Serving with
2303
+ PagedAttention](https://arxiv.org/abs/2309.06180). *SOSP '23*. UC Berkeley.
1845
2304
  - Ollama Team (2024). [Ollama Concurrent Requests and Performance FAQ](https://github.com/ollama/ollama/blob/main/docs/faq.md). `OLLAMA_NUM_PARALLEL` configuration guidance.
1846
2305
 
1847
2306
  ### MCP & Agentic Standards
1848
2307
 
1849
2308
  - Agentic AI Foundation / Linux Foundation (2025). [Model Context Protocol Specification 2025-11-25](https://modelcontextprotocol.io/specification/2025-11-25). (Tool integration protocol)
1850
2309
 
1851
- Full research background, citations, and methodology: [docs/research/](docs/research/)
1852
-
1853
2310
  ---
1854
2311
 
1855
2312
  ## Why AIWG
1856
2313
 
2314
+ AIWG is for people who already use AI assistants and want the work to survive
2315
+ past one conversation. It gives the assistant project-readable instructions,
2316
+ specialist roles, repeatable workflows, and a place to save plans, findings,
2317
+ and decisions.
2318
+
1857
2319
  ### For Individual Developers
1858
2320
 
1859
- **Turn your AI coding assistant from a stateless autocomplete into a project-aware development partner.** Without AIWG, every time your AI assistant restarts you lose all context. With AIWG, the `.aiwg/` directory maintains requirements, architecture decisions, test strategies, and project history across sessions. The agent loop means you can hand off complex multi-step tasks ("migrate this module to TypeScript") and walk away — the agent iterates until completion or escalates when stuck.
2321
+ Use AIWG when a coding task needs project context, not just a one-off answer.
2322
+ The `.aiwg/` directory can hold requirements, architecture notes, test plans,
2323
+ reviews, and task reports. Later sessions can read those artifacts before
2324
+ changing code. Agent-loop workflows can continue a bounded task until the
2325
+ named verification check passes, a limit is reached, or human input is needed.
1860
2326
 
1861
2327
  ### For Engineering Teams
1862
2328
 
1863
- **Standardize how your team works with AI across 14 provider integrations.** Whether your team uses Antigravity, Claude Code, Codex, Copilot, Cursor, Pi, or Warp, everyone can receive the same source workflows adapted to the provider's supported surfaces. The 35 enforcement rules prevent common AI mistakes: deleting tests to make them pass, fabricating citations, hard-coding tokens, or silently dropping features. Human-in-the-loop gates at phase transitions ensure no AI output reaches production without human review.
2329
+ Use AIWG to keep shared instructions and review artifacts with the repository.
2330
+ Teams can deploy the same source workflows to the provider surfaces they use,
2331
+ then verify each provider handoff separately. Human gates and review steps help
2332
+ make important changes visible before they move forward; they do not replace
2333
+ human code review, testing, security review, or release approval.
1864
2334
 
1865
2335
  ### For Platform Engineers
1866
2336
 
1867
- **Deploy consistent AI-augmented workflows across your organization.** AIWG's extension system lets you create custom agents, commands, and skills specific to your domain, then deploy them to any supported platform. The scaffolding commands (`aiwg add-agent`, `aiwg scaffold-addon`, `aiwg scaffold-framework`) make it easy to build and distribute organizational capabilities.
2337
+ Use AIWG when you need reusable AI workflows across projects or teams. The
2338
+ extension system supports project-local rules, skills, agents, addons, and
2339
+ frameworks. Scaffolding commands help create those components, and discovery
2340
+ helps agents find the resulting capabilities without requiring users to
2341
+ memorize the catalog.
1868
2342
 
1869
2343
  ### For Researchers
1870
2344
 
1871
- **Standards-aligned implementation of multi-agent systems with full provenance.** AIWG operationalizes FAIR Principles (G20, EU, NIH endorsement), W3C PROV for provenance tracking, and GRADE methodology for evidence quality. The research framework manages paper discovery, citation integrity, and archival lifecycles. All 138+ research paper citations are grounded in verified sources — no hallucinated references.
2345
+ Use AIWG to organize source notes, citations, provenance records, and research
2346
+ artifacts. The research framework uses ideas from FAIR, W3C PROV, GRADE, and
2347
+ OAIS to make research work easier to inspect and maintain. Those alignments are
2348
+ documentation and workflow structures, not a certification that every generated
2349
+ note is correct.
1872
2350
 
1873
2351
  ### For Security Teams
1874
2352
 
1875
- **Forensics-grade investigation workflows and security review automation.** The Forensics Complete framework provides 13 specialized agents for digital forensics and incident response, with NIST SP 800-86 evidence handling, MITRE ATT&CK mapping, Sigma rule hunting, and STIX 2.1 IOC formatting. Security review cycles integrate into the SDLC with automated threat modeling and vulnerability management.
2353
+ Use AIWG for structured security reviews, incident investigation notes, and
2354
+ forensics-oriented workflows. The forensics framework includes target
2355
+ profiling, triage, acquisition, timeline, IOC, and reporting workflows with
2356
+ references to NIST SP 800-86, MITRE ATT&CK, Sigma, and STIX. Investigations
2357
+ still require authorized access, evidence handling discipline, and human
2358
+ review.
1876
2359
 
1877
2360
  ---
1878
2361
 
1879
2362
  ## AIWG vs Manual AI Workflows
1880
2363
 
1881
- | Capability | Without AIWG | With AIWG |
1882
- |-----------|-------------|-----------|
1883
- | **Context persistence** | Lost on every restart | `.aiwg/` survives across sessions |
1884
- | **Multi-agent coordination** | Manual prompt switching | Orchestrated parallel reviews with synthesis |
1885
- | **Quality enforcement** | Hope for the best | 35 rules auto-enforced (anti-laziness, token security, citation integrity) |
1886
- | **Error recovery** | Start over | Agent loop iterates with learned debug memory |
1887
- | **Long-running tasks** | Babysit the terminal | External agent loop runs 6-8+ hours crash-resilient |
1888
- | **Traceability** | Grep and hope | @-mention system with bidirectional linking |
1889
- | **Reproducibility** | Non-deterministic | Strict mode (temperature=0), checkpoints, validation |
1890
- | **Platform switching** | Rewrite all prompts | `--provider copilot` deploys identical workflows |
1891
- | **Citation integrity** | AI may hallucinate | Retrieval-first architecture, GRADE-assessed sources only |
1892
- | **Phase management** | Ad-hoc | Stage-gate with human approval at transitions |
2364
+ | Task | Manual AI workflow | AIWG mechanism | Limit to keep in view |
2365
+ |------|--------------------|----------------|-----------------------|
2366
+ | Carry context into another session | Paste summaries, links, and decisions again | Save project artifacts under `.aiwg/` and route later work through them | The assistant still has to read and interpret the right artifacts |
2367
+ | Coordinate review perspectives | Ask separate prompts and merge notes by hand | Use specialist roles and review workflows that produce a combined artifact | Reviewers can share wrong assumptions if the input context is wrong |
2368
+ | Keep task scope visible | Track instructions in chat history | Use workflows, rules, and saved task outputs with explicit scope | Prompt-level rules are not hard technical enforcement by themselves |
2369
+ | Recover from a failed attempt | Restart from the last visible message | Use loop state, checkpoints, reports, and bounded retries where configured | Recovery can still stop on missing tools, unclear goals, or failing tests |
2370
+ | Work across AI tools | Rewrite prompts for each provider | Deploy provider-specific files from the same source workflows | Provider capabilities, permissions, and reload behavior differ |
2371
+ | Trace decisions to work products | Search notes manually | Use artifacts, mentions, indexes, and provenance records | Links can drift and need validation |
2372
+ | Review citations and research notes | Trust generated citations or manually inspect each one | Store source records, notes, citations, and quality assessments together | Source-grounding reduces risk; it does not prove every claim is correct |
2373
+ | Move through project phases | Keep phase criteria in a checklist | Use stage-gate workflows with human approval points | Gates reflect configured criteria and available evidence |
1893
2374
 
1894
2375
  ---
1895
2376
 
1896
- ## Standards & Compliance
1897
-
1898
- | Standard | How AIWG Uses It |
1899
- |----------|-----------------|
1900
- | **FAIR Principles** (G20, EU, NIH) | Findable, Accessible, Interoperable, Reusable artifact management |
1901
- | **W3C PROV** | Provenance tracking for all generated artifacts |
1902
- | **GRADE** (WHO, Cochrane, NICE) | Evidence quality assessment for research citations |
1903
- | **OAIS** (ISO 14721) | Archival lifecycle management for research corpus |
1904
- | **NIST SP 800-86** | Digital forensics evidence handling |
1905
- | **MITRE ATT&CK** | Threat technique mapping in forensics framework |
1906
- | **STIX 2.1** | Indicator of Compromise formatting |
1907
- | **Sigma Rules** | Threat detection rule format |
1908
- | **IEEE 830** | Requirements specification traceability |
1909
- | **MCP** (Linux Foundation) | Model Context Protocol for tool integration |
1910
- | **CalVer** | Calendar versioning (YYYY.M.PATCH) |
2377
+ ## Standards Alignment
2378
+
2379
+ AIWG uses standards and established methods as design references. This section
2380
+ describes the intended mapping; it is not a compliance guarantee, audit
2381
+ attestation, or substitute for domain-specific review.
2382
+
2383
+ | Standard or method | How AIWG uses it |
2384
+ |--------------------|------------------|
2385
+ | **FAIR Principles** | Artifact and research-corpus structure that favors findable, accessible, interoperable, and reusable records |
2386
+ | **W3C PROV** | Provenance records for selected generated artifacts and derived outputs |
2387
+ | **GRADE** | Evidence-quality language and review patterns for research citations |
2388
+ | **OAIS** (ISO 14721) | Archival lifecycle concepts for research and media corpus handling |
2389
+ | **NIST SP 800-86** | Digital-forensics evidence-handling references in forensics workflows |
2390
+ | **MITRE ATT&CK** | Threat-technique mapping references for security and forensics analysis |
2391
+ | **STIX 2.1** | Indicator-of-compromise formatting references |
2392
+ | **Sigma Rules** | Threat-detection rule format references |
2393
+ | **IEEE 830** | Requirements-specification and traceability influence for SDLC artifacts |
2394
+ | **MCP** | Model Context Protocol integration for tool-based AI workflows |
2395
+ | **CalVer** | Calendar versioning format for AIWG releases |
1911
2396
 
1912
2397
  ---
1913
2398
 
@@ -1915,25 +2400,27 @@ Full research background, citations, and methodology: [docs/research/](docs/rese
1915
2400
 
1916
2401
  ### Getting Started
1917
2402
 
1918
- - **[Quick Start Guide](docs/quickstart.md)** — Install and deploy in minutes
1919
- - **[Prerequisites](docs/getting-started/prerequisites.md)** — Node.js, AI platforms, OS support
2403
+ - **[Quick Start Guide](docs/quickstart.md)** — Connect a project and get a saved first result
2404
+ - **[Install, Connect, and Verify](docs/getting-started/install-connect-verify.md)** — Canonical first-time setup and
2405
+ repair path
2406
+ - **[Prerequisites](docs/getting-started/prerequisites.md)** — Node.js, AI platforms, and operating-system notes
1920
2407
  - **[Agent and Operator Reference](docs/agents/README.md)** — deterministic
1921
2408
  commands, flags, outputs, and recovery contracts for agents and advanced
1922
2409
  operators
1923
2410
 
1924
2411
  ### Customize
1925
2412
 
1926
- - **[Make AIWG Yours](docs/customization/README.md)** — Personal rules, agents, and skills that go live immediately
1927
- - **[Customization Examples](docs/customization/examples.md)** — 5 concrete examples of what people actually customize
2413
+ - **[Make AIWG Yours](docs/customization/README.md)** — Project-local rules, agents, and skills
2414
+ - **[Customization Examples](docs/customization/examples.md)** — Concrete examples of what teams customize
1928
2415
  - **[Fork Workflow](docs/customization/fork-workflow.md)** — Upstream sync, contributing back, the ownership model
1929
2416
 
1930
2417
  ### By Audience
1931
2418
 
1932
2419
  **Practitioners:**
1933
2420
 
1934
- - [Quick Start Guide](docs/quickstart.md) — Hands-on workflows
1935
- - [Agent Loop Guide](docs/ralph-guide.md) — Iterative execution with crash recovery
1936
- - [Platform Guides](docs/integrations/) — 5-10 minute setup per platform
2421
+ - [Quick Start Guide](docs/quickstart.md) — Hands-on first workflow
2422
+ - [Agent Loop Guide](docs/ralph-guide.md) — Iterative execution with explicit completion checks
2423
+ - [Platform Guides](docs/integrations/) — Provider-specific setup and handoff details
1937
2424
 
1938
2425
  **Technical Leaders:**
1939
2426
 
@@ -1945,27 +2432,36 @@ Full research background, citations, and methodology: [docs/research/](docs/rese
1945
2432
 
1946
2433
  - [Research Background](docs/research/) — Literature review and citations
1947
2434
  - [Glossary](docs/research/glossary.md) — Professional terminology mapping
1948
- - [Production-Grade Guide](docs/production-grade-guide.md) — Failure mode mitigation
2435
+ - [Production-Grade Guide](docs/frameworks/sdlc-complete/production-grade-guide.md) — Failure mode mitigation patterns
1949
2436
 
1950
2437
  ### Platform Guides
1951
2438
 
1952
- - **[Claude Code](docs/integrations/claude-code-quickstart.md)** — 5-10 min setup
1953
- - **[Warp Terminal](docs/integrations/warp-terminal-quickstart.md)** — 3-5 min setup
1954
- - **[Factory AI](docs/integrations/factory-quickstart.md)** — 5-10 min setup
1955
- - **[Cursor](docs/integrations/cursor-quickstart.md)** — 5-10 min setup
1956
- - **[All Integrations](docs/integrations/)**
2439
+ - **[Claude Code](docs/integrations/claude-code-quickstart.md)** — Claude Code setup and handoff
2440
+ - **[OpenAI Codex](docs/integrations/codex-quickstart.md)** — Codex setup and handoff
2441
+ - **[GitHub Copilot](docs/integrations/copilot-quickstart.md)** — Copilot setup and handoff
2442
+ - **[Warp Terminal](docs/integrations/warp-terminal-quickstart.md)** — Warp setup and handoff
2443
+ - **[Factory AI](docs/integrations/factory-quickstart.md)** — Factory setup and handoff
2444
+ - **[Cursor](docs/integrations/cursor-quickstart.md)** — Cursor setup and handoff
2445
+ - **[All Integrations](docs/integrations/)** — Provider guide directory
1957
2446
 
1958
2447
  ### Framework Documentation
1959
2448
 
1960
- - **[SDLC Framework](agentic/code/frameworks/sdlc-complete/README.md)** — 98 agents, phase workflows, quality gates
2449
+ - **[SDLC Framework](agentic/code/frameworks/sdlc-complete/README.md)** — Phase workflows, quality gates, and
2450
+ development artifacts
1961
2451
  - **[Forensics Complete](agentic/code/frameworks/forensics-complete/README.md)** — DFIR investigation workflows
1962
- - **[Marketing Kit](agentic/code/frameworks/media-marketing-kit/README.md)** — 37 agents, campaign lifecycle
2452
+ - **[Marketing Kit](agentic/code/frameworks/media-marketing-kit/README.md)** — Campaign lifecycle, content, brand, and
2453
+ review workflows
1963
2454
  - **[Media Curator](agentic/code/frameworks/media-curator/README.md)** — Media archive management
1964
- - **[Research Complete](agentic/code/frameworks/research-complete/README.md)** — 8-stage research pipeline
2455
+ - **[Research Complete](agentic/code/frameworks/research-complete/README.md)** — Research pipeline, source notes,
2456
+ citation, and archive workflows
2457
+ - **[Knowledge Base](agentic/code/frameworks/knowledge-base/README.md)** — Source ingest, wiki pages, and corpus health
2458
+ - **[Ops Complete](agentic/code/frameworks/ops-complete/README.md)** — Runbooks, infrastructure reviews, and
2459
+ operational workflows
1965
2460
 
1966
2461
  ### Extension System
1967
2462
 
1968
- AIWG's unified extension system enables dynamic discovery, semantic search, and cross-platform deployment:
2463
+ AIWG's extension system supports discovery, semantic search, and
2464
+ cross-platform deployment for project-local and packaged capabilities:
1969
2465
 
1970
2466
  - **[Extension System Overview](docs/extensions/overview.md)** — Architecture and capabilities
1971
2467
  - **[Creating Extensions](docs/extensions/creating-extensions.md)** — Build custom agents, commands, skills
@@ -1974,26 +2470,28 @@ AIWG's unified extension system enables dynamic discovery, semantic search, and
1974
2470
  ### Advanced Topics
1975
2471
 
1976
2472
  - **[Agent Loop](docs/ralph-guide.md)** — Iterative task execution with crash recovery
1977
- - **[RLM Addon](agentic/code/addons/rlm/README.md)** — Recursive context decomposition for 10M+ token processing
1978
- - **[Daemon Mode](docs/daemon-guide.md)** — Background file watching, cron scheduling, IPC
1979
- - **[Messaging Integration](docs/messaging-guide.md)** — Bidirectional Slack, Discord, and Telegram bots
1980
- - **[MCP Server](docs/mcp/)** — Model Context Protocol integration
1981
- - **[Agent Design Bible](docs/AGENT-DESIGN.md)** — 10 Golden Rules for agent creation
2473
+ - **[RLM Addon](agentic/code/addons/rlm/README.md)** — Recursive context decomposition
2474
+ - **[External Automation](docs/getting-started/daemon-and-automation.md)** — Current automation boundaries and
2475
+ external-job contracts
2476
+ - **[MCP Server](docs/mcp/README.md)** — Model Context Protocol integration
2477
+ - **[Agent Design](docs/frameworks/sdlc-complete/agent-design.md)** — Agent creation guidance
1982
2478
  - **[YAML Metalanguage](agentic/code/frameworks/sdlc-complete/schemas/metalanguage/)** — Declarative workflow schemas
2479
+ - **[Usage Notes](docs/usage-notes.md)** — Rate-limit and usage guidance
1983
2480
 
1984
2481
  ---
1985
2482
 
1986
2483
  ## Contributing
1987
2484
 
1988
- We welcome contributions! See [CONTRIBUTING.md](CONTRIBUTING.md) for guidelines.
2485
+ Contributions are welcome. See [CONTRIBUTING.md](CONTRIBUTING.md) for the
2486
+ project guidelines.
1989
2487
 
1990
2488
  **Quick contributions:**
1991
2489
 
1992
- - Found an AI pattern? [Open an issue](https://github.com/jmagly/aiwg/issues/new)
1993
- - Have a better rewrite? Submit a PR to `examples/`
1994
- - Want to add an agent? Use `aiwg add-agent` or see `docs/development/agent-template.md`
1995
- - Want to add a skill? Use `aiwg add-skill`
1996
- - Want to create an addon? Use `aiwg scaffold-addon`
2490
+ - Found a bug or confusing workflow? [Open an issue](https://github.com/jmagly/aiwg/issues/new).
2491
+ - Have a documentation improvement? Submit a PR with the source file and the behavior it clarifies.
2492
+ - Want to add an agent? Use `aiwg add-agent` or see `docs/development/agent-template.md`.
2493
+ - Want to add a skill? Use `aiwg add-skill`.
2494
+ - Want to create an addon? Use `aiwg scaffold-addon`.
1997
2495
 
1998
2496
  ---
1999
2497
 
@@ -2004,13 +2502,16 @@ We welcome contributions! See [CONTRIBUTING.md](CONTRIBUTING.md) for guidelines.
2004
2502
  - **Telegram:** [Join Group](https://t.me/+oJg9w2lE6A5lOGFh)
2005
2503
  - **Issues:** [GitHub Issues](https://github.com/jmagly/aiwg/issues)
2006
2504
  - **Discussions:** [GitHub Discussions](https://github.com/jmagly/aiwg/discussions)
2007
- - **Security:** Report vulnerabilities per [SECURITY.md](SECURITY.md) — do not file public issues.
2505
+ - **Security:** Report vulnerabilities per [SECURITY.md](SECURITY.md); do not file public issues for private
2506
+ vulnerability reports.
2008
2507
 
2009
2508
  ---
2010
2509
 
2011
2510
  ## Badges
2012
2511
 
2013
- Building on AIWG? Show it. Grab a **Built With AIWG** / **Powered By AIWG** badge — light and dark, hosted and free to hot-link (no need to copy the image into your repo):
2512
+ Building on AIWG? You can use a **Built With AIWG** or **Powered By AIWG**
2513
+ badge. Hosted image links are available, and projects can also copy the image
2514
+ if they prefer to avoid hot-linking:
2014
2515
 
2015
2516
  [![Built With AIWG](https://aiwg.io/assets/badges/built-with-aiwg-dark.png)](https://aiwg.io)
2016
2517
 
@@ -2018,24 +2519,30 @@ Building on AIWG? Show it. Grab a **Built With AIWG** / **Powered By AIWG** badg
2018
2519
  [![Built With AIWG](https://aiwg.io/assets/badges/built-with-aiwg-dark.png)](https://aiwg.io)
2019
2520
  ```
2020
2521
 
2021
- Full set with copy-paste snippets at **[aiwg.io/badges](https://aiwg.io/badges)**.
2522
+ Full set with copy-paste snippets: **[aiwg.io/badges](https://aiwg.io/badges)**.
2022
2523
 
2023
2524
  ---
2024
2525
 
2025
2526
  ## Usage Notes
2026
2527
 
2027
- AIWG is optimized for token efficiency. Rules deploy as a consolidated index (~200 lines) instead of 35 individual files (~9,321 lines). Most users on **Claude Pro** or similar plans will have no issues. See [Usage Notes](docs/usage-notes.md) for rate limit guidance.
2528
+ AIWG tries to keep always-loaded context small by using kernel quickrefs,
2529
+ provider-facing indexes, and on-demand discovery. Actual token use depends on
2530
+ the provider, selected workflows, project size, and how much context the agent
2531
+ loads. See [Usage Notes](docs/usage-notes.md) for rate-limit guidance.
2028
2532
 
2029
2533
  ---
2030
2534
 
2031
2535
  ## License
2032
2536
 
2033
- AIWG-authored code is available under the **MIT License**. See [LICENSE](LICENSE).
2537
+ AIWG-authored code is available under the **MIT License**. See
2538
+ [LICENSE](LICENSE).
2034
2539
  Runtime dependencies retain their own licenses; see
2035
2540
  [THIRD_PARTY_NOTICES.md](THIRD_PARTY_NOTICES.md) for the reviewed Fortemi and
2036
2541
  Bytecask AGPL boundary, source links, and inspection instructions.
2037
2542
 
2038
- **Important:** This framework does not provide legal, security, or financial advice. All generated content should be reviewed before use. See [Terms of Use](docs/terms.md) for full disclaimers.
2543
+ This framework does not provide legal, security, financial, medical, or other
2544
+ professional advice. Generated work should be reviewed before use. See
2545
+ [Terms of Use](docs/terms.md) for full terms.
2039
2546
 
2040
2547
  ---
2041
2548
 
@@ -2079,9 +2586,16 @@ Custom AI and blockchain solutions for the digital age.
2079
2586
 
2080
2587
  ## Acknowledgments
2081
2588
 
2082
- **Research foundations:** Built on established principles from cognitive science (Miller 1956, Sweller 1988), multi-agent systems (Jacobs et al. 1991, MetaGPT, AutoGen), software engineering (Cooper 1990, RUP), and recent AI systems research (ReAct, Self-Refine, DSPy, SWE-Agent). Implements standards from FAIR Principles, OAIS (ISO 14721), W3C PROV, GRADE evidence assessment, and MCP protocol (Linux Foundation).
2589
+ **Research foundations:** AIWG draws from cognitive science (Miller 1956,
2590
+ Sweller 1988), multi-agent systems (Jacobs et al. 1991, MetaGPT, AutoGen),
2591
+ software engineering (Cooper 1990, RUP), and AI systems research including
2592
+ ReAct, Self-Refine, DSPy, and SWE-agent. Standards and methods such as FAIR,
2593
+ OAIS, W3C PROV, GRADE, and MCP inform the structure of selected workflows and
2594
+ artifacts.
2083
2595
 
2084
- **Platforms:** Thanks to Anthropic (Claude Code), GitHub (Copilot), Warp, Factory AI, Cursor, and the OpenCode community for building the platforms that enable this work.
2596
+ **Platforms:** Thanks to Anthropic (Claude Code), GitHub (Copilot), Warp,
2597
+ Factory AI, Cursor, OpenCode, and other provider communities for building
2598
+ tools that make project-local AI workflows possible.
2085
2599
 
2086
2600
  ---
2087
2601