aiwg 2026.9.4 → 2026.9.5

This diff represents the content of publicly available package versions that have been released to one of the supported registries. The information contained in this diff is provided for informational purposes only and reflects changes between package versions as they appear in their respective public registries.
Files changed (524) hide show
  1. package/README.md +1412 -900
  2. package/THIRD_PARTY_NOTICES.md +16 -0
  3. package/agentic/code/addons/agent-loop/README.md +2 -0
  4. package/agentic/code/addons/agent-loop/docs/quickstart.md +18 -13
  5. package/agentic/code/addons/aiwg-dev/docs/overview.md +11 -2
  6. package/agentic/code/addons/aiwg-dev/docs/quickstart.md +8 -4
  7. package/agentic/code/addons/aiwg-utils/agents/aiwg-finder.md +4 -0
  8. package/agentic/code/addons/aiwg-utils/agents/session-analyst.md +44 -0
  9. package/agentic/code/addons/aiwg-utils/commands/sessions.md +33 -0
  10. package/agentic/code/addons/aiwg-utils/docs/overview.md +16 -6
  11. package/agentic/code/addons/aiwg-utils/flows/capabilities/session-evidence-collect.yaml +13 -0
  12. package/agentic/code/addons/aiwg-utils/flows/capabilities/session-evidence-synthesize.yaml +13 -0
  13. package/agentic/code/addons/aiwg-utils/flows/session-investigation.yaml +17 -0
  14. package/agentic/code/addons/aiwg-utils/manifest.json +16 -4
  15. package/agentic/code/addons/aiwg-utils/skills/aiwg-language-map/SKILL.md +10 -0
  16. package/agentic/code/addons/aiwg-utils/skills/cost-history/SKILL.md +5 -0
  17. package/agentic/code/addons/aiwg-utils/skills/session/SKILL.md +3 -0
  18. package/agentic/code/addons/aiwg-utils/skills/session-explore/SKILL.md +105 -0
  19. package/agentic/code/addons/aiwg-utils/skills/session-explore/references/recipes.md +146 -0
  20. package/agentic/code/addons/aiwg-utils/skills/session-harvest/SKILL.md +73 -0
  21. package/agentic/code/addons/aiwg-utils/skills/summarize-transcript/SKILL.md +9 -0
  22. package/agentic/code/addons/auto-memory/docs/overview.md +23 -4
  23. package/agentic/code/addons/civic-action/docs/overview.md +8 -1
  24. package/agentic/code/addons/civic-action/docs/quickstart.md +3 -1
  25. package/agentic/code/addons/compound-memory/docs/overview.md +12 -2
  26. package/agentic/code/addons/daemon/docs/quickstart.md +3 -1
  27. package/agentic/code/addons/dataset-intelligence/docs/overview.md +10 -3
  28. package/agentic/code/addons/dataset-intelligence/skills/dataset-intake/SKILL.md +6 -0
  29. package/agentic/code/addons/guided-implementation/docs/overview.md +22 -3
  30. package/agentic/code/addons/line-memory/docs/overview.md +7 -4
  31. package/agentic/code/addons/network-analysis/CHANGELOG.md +18 -0
  32. package/agentic/code/addons/network-analysis/README.md +108 -0
  33. package/agentic/code/addons/network-analysis/docs/integrations.md +20 -0
  34. package/agentic/code/addons/network-analysis/docs/maintainer-guide.md +105 -0
  35. package/agentic/code/addons/network-analysis/docs/offline-analysis.md +80 -0
  36. package/agentic/code/addons/network-analysis/docs/operator-guide.md +119 -0
  37. package/agentic/code/addons/network-analysis/docs/overview.md +30 -0
  38. package/agentic/code/addons/network-analysis/docs/release-checklist.md +48 -0
  39. package/agentic/code/addons/network-analysis/docs/termshark-handoff.md +72 -0
  40. package/agentic/code/addons/network-analysis/manifest.json +78 -0
  41. package/agentic/code/addons/network-analysis/recipes/README.md +42 -0
  42. package/agentic/code/addons/network-analysis/recipes/beaconing-timing.json +15 -0
  43. package/agentic/code/addons/network-analysis/recipes/before-after.json +15 -0
  44. package/agentic/code/addons/network-analysis/recipes/dns.json +15 -0
  45. package/agentic/code/addons/network-analysis/recipes/endpoints-conversations.json +15 -0
  46. package/agentic/code/addons/network-analysis/recipes/http-metadata.json +15 -0
  47. package/agentic/code/addons/network-analysis/recipes/overview.json +27 -0
  48. package/agentic/code/addons/network-analysis/recipes/stream-selection.json +15 -0
  49. package/agentic/code/addons/network-analysis/recipes/tcp-health.json +15 -0
  50. package/agentic/code/addons/network-analysis/recipes/tls.json +15 -0
  51. package/agentic/code/addons/network-analysis/rules/RULES-INDEX.md +5 -0
  52. package/agentic/code/addons/network-analysis/rules/network-analysis-safety.md +29 -0
  53. package/agentic/code/addons/network-analysis/schemas/network-analysis-contracts.md +19 -0
  54. package/agentic/code/addons/network-analysis/skills/analyze-network-capture/SKILL.md +65 -0
  55. package/agentic/code/addons/network-analysis/templates/analysis-request.md +18 -0
  56. package/agentic/code/addons/network-analysis/templates/termshark-handoff.md +19 -0
  57. package/agentic/code/addons/prose-integration/docs/overview.md +16 -4
  58. package/agentic/code/addons/semantic-memory/skills/memory-ingest/SKILL.md +9 -0
  59. package/agentic/code/addons/testing-quality/docs/overview.md +18 -6
  60. package/agentic/code/addons/testing-quality/docs/quickstart.md +23 -15
  61. package/agentic/code/addons/voice-framework/README.md +16 -1
  62. package/agentic/code/addons/voice-framework/docs/contextual-diagnostics-evaluation.md +52 -0
  63. package/agentic/code/addons/voice-framework/docs/contextual-diagnostics.md +66 -0
  64. package/agentic/code/addons/voice-framework/docs/natural-voice/ADR-001-evidence-and-ownership.md +29 -0
  65. package/agentic/code/addons/voice-framework/docs/natural-voice/evidence-ledger.v1.json +434 -0
  66. package/agentic/code/addons/voice-framework/docs/natural-voice/validate-ledger.mjs +52 -0
  67. package/agentic/code/addons/voice-framework/docs/natural-voice/verification-2026-09-06.md +19 -0
  68. package/agentic/code/addons/voice-framework/docs/overview.md +32 -7
  69. package/agentic/code/addons/voice-framework/docs/quickstart.md +25 -9
  70. package/agentic/code/addons/voice-framework/docs/voice-output-impact.md +54 -0
  71. package/agentic/code/addons/voice-framework/docs/writing-workflows.md +53 -0
  72. package/agentic/code/addons/voice-framework/flows/voice-critique-correction.flow.yaml +595 -0
  73. package/agentic/code/addons/voice-framework/manifest.json +8 -4
  74. package/agentic/code/addons/voice-framework/skills/manifest.json +53 -12
  75. package/agentic/code/addons/voice-framework/skills/output-mode-guide/SKILL.md +12 -0
  76. package/agentic/code/addons/voice-framework/skills/voice-analyze/SKILL.md +18 -0
  77. package/agentic/code/addons/voice-framework/skills/voice-apply/SKILL.md +47 -7
  78. package/agentic/code/addons/voice-framework/skills/voice-blend/SKILL.md +8 -0
  79. package/agentic/code/addons/voice-framework/skills/voice-create/SKILL.md +19 -0
  80. package/agentic/code/addons/writing-quality/README.md +8 -2
  81. package/agentic/code/addons/writing-quality/agents/content-diversifier.md +31 -28
  82. package/agentic/code/addons/writing-quality/agents/prompt-optimizer.md +25 -267
  83. package/agentic/code/addons/writing-quality/agents/writing-validator.md +28 -34
  84. package/agentic/code/addons/writing-quality/context/quick-reference.md +4 -4
  85. package/agentic/code/addons/writing-quality/core/philosophy.md +6 -6
  86. package/agentic/code/addons/writing-quality/manifest.json +1 -1
  87. package/agentic/code/addons/writing-quality/skills/ai-pattern-detection/SKILL.md +20 -16
  88. package/agentic/code/addons/writing-quality/skills/ai-pattern-detection/references/quick-patterns.md +6 -4
  89. package/agentic/code/addons/writing-quality/skills/ai-pattern-detection/scripts/pattern_scanner.py +13 -5
  90. package/agentic/code/addons/writing-quality/validation/banned-patterns.md +5 -7
  91. package/agentic/code/addons/writing-quality/validation/scoring-config.json +9 -9
  92. package/agentic/code/addons/writing-quality/validation/validation-checklist.md +22 -26
  93. package/agentic/code/frameworks/forensics-complete/README.md +4 -0
  94. package/agentic/code/frameworks/forensics-complete/agents/network-analyst.md +6 -0
  95. package/agentic/code/frameworks/forensics-complete/docs/packet-evidence-integration.md +78 -0
  96. package/agentic/code/frameworks/forensics-complete/templates/chain-of-custody.md +13 -0
  97. package/agentic/code/frameworks/knowledge-base/docs/overview.md +20 -2
  98. package/agentic/code/frameworks/media-curator/docs/overview.md +17 -6
  99. package/agentic/code/frameworks/media-marketing-kit/agents/content-writer.md +16 -7
  100. package/agentic/code/frameworks/media-marketing-kit/agents/copywriter.md +9 -0
  101. package/agentic/code/frameworks/media-marketing-kit/agents/email-marketer.md +9 -0
  102. package/agentic/code/frameworks/media-marketing-kit/agents/social-media-specialist.md +9 -0
  103. package/agentic/code/frameworks/media-marketing-kit/docs/overview.md +23 -15
  104. package/agentic/code/frameworks/media-marketing-kit/docs/quickstart.md +28 -8
  105. package/agentic/code/frameworks/media-marketing-kit/skills/pr-launch/SKILL.md +9 -0
  106. package/agentic/code/frameworks/media-marketing-kit/templates/content/blog-post-template.md +5 -5
  107. package/agentic/code/frameworks/media-marketing-kit/templates/social/social-post-template.md +6 -0
  108. package/agentic/code/frameworks/ops-complete/README.md +1 -0
  109. package/agentic/code/frameworks/ops-complete/docs/overview.md +17 -7
  110. package/agentic/code/frameworks/ops-complete/docs/packet-verification.md +53 -0
  111. package/agentic/code/frameworks/ops-complete/docs/quickstart.md +23 -3
  112. package/agentic/code/frameworks/research-complete/README.md +1 -0
  113. package/agentic/code/frameworks/research-complete/config/source-types.yaml +11 -0
  114. package/agentic/code/frameworks/research-complete/docs/overview.md +39 -26
  115. package/agentic/code/frameworks/research-complete/docs/packet-evidence.md +62 -0
  116. package/agentic/code/frameworks/research-complete/docs/quickstart.md +30 -12
  117. package/agentic/code/frameworks/research-complete/skills/induct-research/SKILL.md +2 -0
  118. package/agentic/code/frameworks/research-complete/skills/research-acquire/SKILL.md +2 -0
  119. package/agentic/code/frameworks/research-complete/skills/research-workflow/SKILL.md +2 -0
  120. package/agentic/code/frameworks/research-complete/skills/source-types/SKILL.md +1 -1
  121. package/agentic/code/frameworks/research-complete/templates/manifest.json +14 -3
  122. package/agentic/code/frameworks/research-complete/templates/reference-packet-evidence.md +75 -0
  123. package/agentic/code/frameworks/sdlc-complete/schemas/verification-contract.schema.json +1 -0
  124. package/agentic/code/frameworks/sdlc-complete/skills/issue-planner/SKILL.md +19 -0
  125. package/agentic/code/frameworks/sdlc-complete/templates/deepseek-harness/AGENTS.md.aiwg-template +20 -0
  126. package/agentic/code/frameworks/sdlc-complete/templates/test/manifest.json +2 -0
  127. package/agentic/code/frameworks/sdlc-complete/templates/test/packet-evidence-defect.md +19 -0
  128. package/agentic/code/frameworks/sdlc-complete/templates/test/packet-evidence-test-plan.md +27 -0
  129. package/agentic/code/frameworks/security-engineering/docs/network-control-review.md +21 -0
  130. package/agentic/code/plugins/agent-loop/.claude-plugin/plugin.json +1 -1
  131. package/agentic/code/plugins/agent-loop/README.md +2 -0
  132. package/agentic/code/plugins/agent-loop/docs/quickstart.md +18 -13
  133. package/agentic/code/plugins/agent-persistence/.claude-plugin/plugin.json +1 -1
  134. package/agentic/code/plugins/agentic-installer/.claude-plugin/plugin.json +1 -1
  135. package/agentic/code/plugins/aiwg-dev/.claude-plugin/plugin.json +1 -1
  136. package/agentic/code/plugins/aiwg-dev/docs/overview.md +11 -2
  137. package/agentic/code/plugins/aiwg-dev/docs/quickstart.md +8 -4
  138. package/agentic/code/plugins/aiwg-evals/.claude-plugin/plugin.json +1 -1
  139. package/agentic/code/plugins/auto-memory/.claude-plugin/plugin.json +1 -1
  140. package/agentic/code/plugins/auto-memory/docs/overview.md +23 -4
  141. package/agentic/code/plugins/browser-control/.claude-plugin/plugin.json +1 -1
  142. package/agentic/code/plugins/color-palette/.claude-plugin/plugin.json +1 -1
  143. package/agentic/code/plugins/compound-memory/.claude-plugin/plugin.json +1 -1
  144. package/agentic/code/plugins/compound-memory/docs/overview.md +12 -2
  145. package/agentic/code/plugins/context-curator/.claude-plugin/plugin.json +1 -1
  146. package/agentic/code/plugins/daemon/.claude-plugin/plugin.json +1 -1
  147. package/agentic/code/plugins/daemon/docs/quickstart.md +3 -1
  148. package/agentic/code/plugins/doc-intelligence/.claude-plugin/plugin.json +1 -1
  149. package/agentic/code/plugins/droid-bridge/.claude-plugin/plugin.json +1 -1
  150. package/agentic/code/plugins/forensics/.claude-plugin/plugin.json +1 -1
  151. package/agentic/code/plugins/forensics/agents/network-analyst.md +6 -0
  152. package/agentic/code/plugins/guided-implementation/.claude-plugin/plugin.json +1 -1
  153. package/agentic/code/plugins/guided-implementation/docs/overview.md +22 -3
  154. package/agentic/code/plugins/hooks/.claude-plugin/plugin.json +1 -1
  155. package/agentic/code/plugins/knowledge-base/.claude-plugin/plugin.json +1 -1
  156. package/agentic/code/plugins/line-memory/.claude-plugin/plugin.json +1 -1
  157. package/agentic/code/plugins/line-memory/docs/overview.md +7 -4
  158. package/agentic/code/plugins/llm-wiki/.claude-plugin/plugin.json +1 -1
  159. package/agentic/code/plugins/marketing/.claude-plugin/plugin.json +1 -1
  160. package/agentic/code/plugins/marketing/agents/content-writer.md +16 -7
  161. package/agentic/code/plugins/marketing/agents/copywriter.md +9 -0
  162. package/agentic/code/plugins/marketing/agents/email-marketer.md +9 -0
  163. package/agentic/code/plugins/marketing/agents/social-media-specialist.md +9 -0
  164. package/agentic/code/plugins/marketing/skills/pr-launch/SKILL.md +9 -0
  165. package/agentic/code/plugins/media-curator/.claude-plugin/plugin.json +1 -1
  166. package/agentic/code/plugins/nlp-prod/.claude-plugin/plugin.json +1 -1
  167. package/agentic/code/plugins/ops/.claude-plugin/plugin.json +1 -1
  168. package/agentic/code/plugins/prose-integration/.claude-plugin/plugin.json +1 -1
  169. package/agentic/code/plugins/prose-integration/docs/overview.md +16 -4
  170. package/agentic/code/plugins/research/.claude-plugin/plugin.json +1 -1
  171. package/agentic/code/plugins/research/skills/induct-research/SKILL.md +2 -0
  172. package/agentic/code/plugins/research/skills/research-acquire/SKILL.md +2 -0
  173. package/agentic/code/plugins/research/skills/research-workflow/SKILL.md +2 -0
  174. package/agentic/code/plugins/research/skills/source-types/SKILL.md +1 -1
  175. package/agentic/code/plugins/rlm/.claude-plugin/plugin.json +1 -1
  176. package/agentic/code/plugins/sdlc/.claude-plugin/plugin.json +1 -1
  177. package/agentic/code/plugins/sdlc/skills/issue-planner/SKILL.md +19 -0
  178. package/agentic/code/plugins/security-engineering/.claude-plugin/plugin.json +1 -1
  179. package/agentic/code/plugins/semantic-memory/.claude-plugin/plugin.json +1 -1
  180. package/agentic/code/plugins/semantic-memory/skills/memory-ingest/SKILL.md +9 -0
  181. package/agentic/code/plugins/skill-factory/.claude-plugin/plugin.json +1 -1
  182. package/agentic/code/plugins/star-prompt/.claude-plugin/plugin.json +1 -1
  183. package/agentic/code/plugins/testing-quality/.claude-plugin/plugin.json +1 -1
  184. package/agentic/code/plugins/testing-quality/docs/overview.md +18 -6
  185. package/agentic/code/plugins/testing-quality/docs/quickstart.md +23 -15
  186. package/agentic/code/plugins/twelve-factor/.claude-plugin/plugin.json +1 -1
  187. package/agentic/code/plugins/uat-mcp/.claude-plugin/plugin.json +1 -1
  188. package/agentic/code/plugins/utils/.claude-plugin/plugin.json +1 -1
  189. package/agentic/code/plugins/utils/agents/aiwg-finder.md +4 -0
  190. package/agentic/code/plugins/utils/agents/session-analyst.md +44 -0
  191. package/agentic/code/plugins/utils/skills/aiwg-language-map/SKILL.md +10 -0
  192. package/agentic/code/plugins/utils/skills/cost-history/SKILL.md +5 -0
  193. package/agentic/code/plugins/utils/skills/session/SKILL.md +3 -0
  194. package/agentic/code/plugins/utils/skills/session-explore/SKILL.md +105 -0
  195. package/agentic/code/plugins/utils/skills/session-explore/references/recipes.md +146 -0
  196. package/agentic/code/plugins/utils/skills/session-harvest/SKILL.md +73 -0
  197. package/agentic/code/plugins/utils/skills/summarize-transcript/SKILL.md +9 -0
  198. package/agentic/code/plugins/validation-complete/.claude-plugin/plugin.json +1 -1
  199. package/agentic/code/plugins/verbalized-sampling/.claude-plugin/plugin.json +1 -1
  200. package/agentic/code/plugins/voice/.claude-plugin/plugin.json +1 -1
  201. package/agentic/code/plugins/voice/skills/manifest.json +53 -12
  202. package/agentic/code/plugins/voice/skills/output-mode-guide/SKILL.md +12 -0
  203. package/agentic/code/plugins/voice/skills/voice-analyze/SKILL.md +18 -0
  204. package/agentic/code/plugins/voice/skills/voice-apply/SKILL.md +47 -7
  205. package/agentic/code/plugins/voice/skills/voice-blend/SKILL.md +8 -0
  206. package/agentic/code/plugins/voice/skills/voice-create/SKILL.md +19 -0
  207. package/agentic/code/plugins/writing/.claude-plugin/plugin.json +1 -1
  208. package/agentic/code/plugins/writing/agents/content-diversifier.md +31 -28
  209. package/agentic/code/plugins/writing/agents/prompt-optimizer.md +25 -267
  210. package/agentic/code/plugins/writing/agents/writing-validator.md +28 -34
  211. package/agentic/code/plugins/writing/skills/ai-pattern-detection/SKILL.md +20 -16
  212. package/agentic/code/plugins/writing/skills/ai-pattern-detection/references/quick-patterns.md +6 -4
  213. package/agentic/code/plugins/writing/skills/ai-pattern-detection/scripts/pattern_scanner.py +13 -5
  214. package/agentic/code/providers/capability-matrix.yaml +45 -3
  215. package/agentic/code/providers/deepseek-harness/README.md +8 -0
  216. package/agentic/code/providers/deepseek-harness/aiwg.cordis.patch.yml +10 -0
  217. package/agentic/code/providers/model-capabilities.v1.json +11 -0
  218. package/agentic/code/providers/model-catalog.v1.json +8 -0
  219. package/bin/aiwg.mjs +1 -0
  220. package/dist/src/api/index.d.ts +14 -0
  221. package/dist/src/api/index.js +14 -0
  222. package/dist/src/artifacts/corpus-tools/source-types.js +1 -0
  223. package/dist/src/artifacts/index-files.js +17 -2
  224. package/dist/src/artifacts/repair.js +55 -6
  225. package/dist/src/catalog/cli.js +21 -7
  226. package/dist/src/catalog/cli.mjs +22 -7
  227. package/dist/src/cli/agent-spawn.js +10 -1
  228. package/dist/src/cli/handlers/artifacts.js +22 -3
  229. package/dist/src/cli/handlers/help.js +3 -0
  230. package/dist/src/cli/handlers/index.js +5 -1
  231. package/dist/src/cli/handlers/models.js +2 -2
  232. package/dist/src/cli/handlers/output-mode.js +1 -1
  233. package/dist/src/cli/handlers/runtime-info.js +1 -1
  234. package/dist/src/cli/handlers/sessions.js +55 -15
  235. package/dist/src/cli/handlers/steward.js +1 -1
  236. package/dist/src/cli/handlers/subcommands.js +5 -0
  237. package/dist/src/cli/handlers/writer-profile.js +110 -0
  238. package/dist/src/cli/handlers/writing.js +122 -0
  239. package/dist/src/cli/router.js +5 -1
  240. package/dist/src/config/project-artifacts-runtime.mjs +33 -1
  241. package/dist/src/config/project-artifacts.js +1 -1
  242. package/dist/src/dataset/fortemi-dataset-execution.d.ts +23 -0
  243. package/dist/src/dataset/fortemi-dataset-execution.js +158 -0
  244. package/dist/src/dataset/fortemi-live-qualification.d.ts +4 -2
  245. package/dist/src/dataset/fortemi-live-qualification.js +20 -21
  246. package/dist/src/dataset/fortemi-run-receipt.d.ts +44 -0
  247. package/dist/src/dataset/fortemi-run-receipt.js +74 -0
  248. package/dist/src/dataset/index.d.ts +2 -0
  249. package/dist/src/dataset/index.js +2 -0
  250. package/dist/src/extensions/commands/definitions.js +24 -0
  251. package/dist/src/extensions/manifest.js +3 -0
  252. package/dist/src/mcp/server.mjs +2 -0
  253. package/dist/src/mcp/tools/writer-profiles.mjs +40 -0
  254. package/dist/src/models/provider-policy.js +1 -1
  255. package/dist/src/network-analysis/analyzer.d.ts +80 -0
  256. package/dist/src/network-analysis/analyzer.js +667 -0
  257. package/dist/src/network-analysis/citations.d.ts +76 -0
  258. package/dist/src/network-analysis/citations.js +107 -0
  259. package/dist/src/network-analysis/forensics.d.ts +131 -0
  260. package/dist/src/network-analysis/forensics.js +132 -0
  261. package/dist/src/network-analysis/governance.d.ts +94 -0
  262. package/dist/src/network-analysis/governance.js +216 -0
  263. package/dist/src/network-analysis/index.d.ts +10 -0
  264. package/dist/src/network-analysis/index.js +10 -0
  265. package/dist/src/network-analysis/probe.d.ts +90 -0
  266. package/dist/src/network-analysis/probe.js +405 -0
  267. package/dist/src/network-analysis/recipes.d.ts +102 -0
  268. package/dist/src/network-analysis/recipes.js +88 -0
  269. package/dist/src/network-analysis/research.d.ts +119 -0
  270. package/dist/src/network-analysis/research.js +181 -0
  271. package/dist/src/network-analysis/termshark.d.ts +95 -0
  272. package/dist/src/network-analysis/termshark.js +252 -0
  273. package/dist/src/network-analysis/verification.d.ts +106 -0
  274. package/dist/src/network-analysis/verification.js +171 -0
  275. package/dist/src/output-modes/registry.js +37 -6
  276. package/dist/src/output-modes/runtime.js +164 -28
  277. package/dist/src/providers/provider-definitions.js +49 -0
  278. package/dist/src/providers/provider-inventory.js +1 -0
  279. package/dist/src/providers/transformation-receipt.js +3 -2
  280. package/dist/src/sessions/adapters/deepseek-harness.js +178 -0
  281. package/dist/src/sessions/batch-import.js +7 -0
  282. package/dist/src/sessions/contracts.js +2 -1
  283. package/dist/src/sessions/index.js +1 -0
  284. package/dist/src/sessions/workspace-discovery.js +5 -1
  285. package/dist/src/skills/deployer.js +6 -6
  286. package/dist/src/smiths/context-pipeline/workspace-context.js +7 -0
  287. package/dist/src/writing/channel-packs.d.ts +9 -0
  288. package/dist/src/writing/channel-packs.js +13 -0
  289. package/dist/src/writing/contextual-diagnostics.d.ts +66 -0
  290. package/dist/src/writing/contextual-diagnostics.js +142 -0
  291. package/dist/src/writing/example-generator.js +7 -6
  292. package/dist/src/writing/exemplar-selection.d.ts +164 -0
  293. package/dist/src/writing/exemplar-selection.js +186 -0
  294. package/dist/src/writing/fidelity.d.ts +22 -0
  295. package/dist/src/writing/fidelity.js +61 -0
  296. package/dist/src/writing/validation-engine.js +32 -15
  297. package/dist/src/writing/voice-evaluation.d.ts +547 -0
  298. package/dist/src/writing/voice-evaluation.js +301 -0
  299. package/dist/src/writing/voice-revision.d.ts +170 -0
  300. package/dist/src/writing/voice-revision.js +201 -0
  301. package/dist/src/writing/writer-migration.d.ts +124 -0
  302. package/dist/src/writing/writer-migration.js +216 -0
  303. package/dist/src/writing/writer-profile-legacy.d.ts +12 -0
  304. package/dist/src/writing/writer-profile-legacy.js +145 -0
  305. package/dist/src/writing/writer-profile-store.d.ts +24 -0
  306. package/dist/src/writing/writer-profile-store.js +117 -0
  307. package/dist/src/writing/writer-profile.d.ts +605 -0
  308. package/dist/src/writing/writer-profile.js +222 -0
  309. package/dist/src/writing/writing-brief.d.ts +423 -0
  310. package/dist/src/writing/writing-brief.js +166 -0
  311. package/dist/src/writing/writing-channels.d.ts +46 -0
  312. package/dist/src/writing/writing-channels.js +63 -0
  313. package/dist/src/writing/writing-consumer.d.ts +36 -0
  314. package/dist/src/writing/writing-consumer.js +39 -0
  315. package/dist/src/writing/writing-receipt.d.ts +1569 -0
  316. package/dist/src/writing/writing-receipt.js +266 -0
  317. package/docs/_manifest.json +210 -20
  318. package/docs/addons/agent-loop/quickstart.md +18 -13
  319. package/docs/addons/aiwg-dev/overview.md +11 -2
  320. package/docs/addons/aiwg-dev/quickstart.md +8 -4
  321. package/docs/addons/aiwg-utils/overview.md +16 -6
  322. package/docs/addons/auto-memory/overview.md +23 -4
  323. package/docs/addons/civic-action/overview.md +8 -1
  324. package/docs/addons/civic-action/quickstart.md +3 -1
  325. package/docs/addons/compound-memory/overview.md +12 -2
  326. package/docs/addons/daemon/quickstart.md +3 -1
  327. package/docs/addons/dataset-intelligence/overview.md +10 -3
  328. package/docs/addons/guided-implementation/overview.md +22 -3
  329. package/docs/addons/line-memory/overview.md +7 -4
  330. package/docs/addons/network-analysis/integrations.md +20 -0
  331. package/docs/addons/network-analysis/maintainer-guide.md +105 -0
  332. package/docs/addons/network-analysis/offline-analysis.md +80 -0
  333. package/docs/addons/network-analysis/operator-guide.md +119 -0
  334. package/docs/addons/network-analysis/overview.md +30 -0
  335. package/docs/addons/network-analysis/release-checklist.md +48 -0
  336. package/docs/addons/network-analysis/termshark-handoff.md +72 -0
  337. package/docs/addons/prose-integration/overview.md +16 -4
  338. package/docs/addons/ralph/quickstart.md +19 -13
  339. package/docs/addons/rlm/deployment-guide.md +2 -2
  340. package/docs/addons/rlm/multi-provider-guide.md +1 -1
  341. package/docs/addons/testing-quality/overview.md +18 -6
  342. package/docs/addons/testing-quality/quickstart.md +23 -15
  343. package/docs/addons/voice-framework/overview.md +32 -7
  344. package/docs/addons/voice-framework/quickstart.md +25 -9
  345. package/docs/architecture/adr-deepseek-harness-runtime.md +70 -0
  346. package/docs/architecture/network-analysis.md +147 -0
  347. package/docs/architecture/schema-inventory.md +8 -0
  348. package/docs/architecture-overview.md +70 -96
  349. package/docs/autonomy-and-human-roles.md +25 -0
  350. package/docs/cli/agent-usage.md +5 -3
  351. package/docs/cli/capability-routing.md +2 -2
  352. package/docs/cli/reference.md +13 -5
  353. package/docs/cockpit/README.md +4 -2
  354. package/docs/config.json +33 -23
  355. package/docs/context-engineering-contract.md +26 -0
  356. package/docs/context-management-patterns.md +2 -0
  357. package/docs/customization/README.md +8 -7
  358. package/docs/customization/project-local-quickstart.md +22 -0
  359. package/docs/development/architecture-illustrations.md +22 -0
  360. package/docs/development/cli-help-routing-audit.md +52 -0
  361. package/docs/development/devkit-overview.md +14 -13
  362. package/docs/development/skill-inventory.md +2 -2
  363. package/docs/docs-sources.json +27 -0
  364. package/docs/extensions/overview.md +11 -7
  365. package/docs/frameworks/forensics-complete/packet-evidence-integration.md +78 -0
  366. package/docs/frameworks/knowledge-base/overview.md +20 -2
  367. package/docs/frameworks/media-curator/overview.md +17 -6
  368. package/docs/frameworks/media-marketing-kit/overview.md +23 -15
  369. package/docs/frameworks/media-marketing-kit/quickstart.md +28 -8
  370. package/docs/frameworks/ops-complete/overview.md +17 -7
  371. package/docs/frameworks/ops-complete/packet-verification.md +53 -0
  372. package/docs/frameworks/ops-complete/quickstart.md +23 -3
  373. package/docs/frameworks/research-complete/overview.md +39 -26
  374. package/docs/frameworks/research-complete/packet-evidence.md +62 -0
  375. package/docs/frameworks/research-complete/quickstart.md +30 -12
  376. package/docs/frameworks/security-engineering/network-control-review.md +21 -0
  377. package/docs/getting-started/README.md +49 -116
  378. package/docs/getting-started/audit-existing-code.md +20 -21
  379. package/docs/getting-started/cognitive-walkthrough-1385-1386.md +2 -1
  380. package/docs/getting-started/daemon-and-automation.md +2 -1
  381. package/docs/getting-started/demo-script.md +2 -1
  382. package/docs/getting-started/existing-project.md +14 -38
  383. package/docs/getting-started/first-success-ask-steward.md +2 -1
  384. package/docs/getting-started/first-success-find-capability.md +2 -1
  385. package/docs/getting-started/first-success-start-intake.md +2 -1
  386. package/docs/getting-started/flow-and-gate-process.md +10 -5
  387. package/docs/getting-started/forensics-framework.md +13 -9
  388. package/docs/getting-started/install-connect-verify.md +16 -9
  389. package/docs/getting-started/just-try-it.md +40 -39
  390. package/docs/getting-started/key-addons.md +69 -19
  391. package/docs/getting-started/language-map.md +2 -1
  392. package/docs/getting-started/macos-install.md +2 -1
  393. package/docs/getting-started/marketing-framework.md +18 -11
  394. package/docs/getting-started/media-curator-framework.md +24 -12
  395. package/docs/getting-started/new-project.md +15 -28
  396. package/docs/getting-started/onboarding-research-refresh.md +2 -1
  397. package/docs/getting-started/onboarding-validation.md +2 -1
  398. package/docs/getting-started/prerequisites.md +17 -14
  399. package/docs/getting-started/provider-handoff.md +3 -3
  400. package/docs/getting-started/research-framework.md +24 -13
  401. package/docs/getting-started/scope-and-recovery.md +2 -1
  402. package/docs/getting-started/sdlc-framework.md +23 -14
  403. package/docs/getting-started/session-history.md +13 -0
  404. package/docs/getting-started/share-aiwg.md +2 -1
  405. package/docs/getting-started/start-here.md +47 -83
  406. package/docs/getting-started/storage-and-pkm.md +2 -1
  407. package/docs/getting-started/team-setup.md +34 -18
  408. package/docs/getting-started/verify-aiwg-is-working.md +12 -9
  409. package/docs/getting-started/writing-and-content.md +14 -10
  410. package/docs/how-it-works.md +86 -373
  411. package/docs/integrations/_manifest.json +1 -0
  412. package/docs/integrations/claude-code-quickstart.md +16 -13
  413. package/docs/integrations/codex-quickstart.md +17 -14
  414. package/docs/integrations/copilot-quickstart.md +17 -9
  415. package/docs/integrations/cross-platform-overview.md +43 -17
  416. package/docs/integrations/cursor-quickstart.md +17 -8
  417. package/docs/integrations/deepseek-harness-quickstart.md +40 -0
  418. package/docs/integrations/factory-quickstart.md +17 -8
  419. package/docs/integrations/hermes-quickstart.md +17 -8
  420. package/docs/integrations/openclaw-quickstart.md +17 -8
  421. package/docs/integrations/opencode-quickstart.md +17 -8
  422. package/docs/integrations/openhuman-quickstart.md +17 -8
  423. package/docs/integrations/pi-quickstart.md +15 -0
  424. package/docs/integrations/warp-terminal-quickstart.md +17 -8
  425. package/docs/integrations/windsurf-quickstart.md +18 -5
  426. package/docs/mcp/README.md +5 -2
  427. package/docs/models/hybrid-architectures.md +2 -0
  428. package/docs/network-analysis/compatibility.md +86 -0
  429. package/docs/overview/capabilities.md +61 -0
  430. package/docs/overview/executive-brief.md +48 -199
  431. package/docs/overview/reading-list.md +259 -0
  432. package/docs/overview/what-is-aiwg.md +71 -679
  433. package/docs/planning/session-intelligence/provider-conformance-matrix.json +11 -1
  434. package/docs/project-local/overview.md +9 -4
  435. package/docs/providers/deepseek-harness-sessions.md +30 -0
  436. package/docs/providers/deepseek-harness.md +197 -0
  437. package/docs/providers/indexed-access-audit.md +41 -0
  438. package/docs/providers/marketplace-consumer.md +2 -2
  439. package/docs/providers/provider-inventory.md +7 -3
  440. package/docs/quickstart-mmk.md +4 -3
  441. package/docs/quickstart-sdlc.md +3 -2
  442. package/docs/quickstart.md +28 -47
  443. package/docs/releases/v2026.9.5-announcement.md +97 -0
  444. package/docs/security/network-analysis-construction-gate.md +46 -0
  445. package/docs/security/network-analysis-threat-model.md +135 -0
  446. package/docs/sessions/cli.md +5 -1
  447. package/docs/sessions/exploration-validation.md +42 -0
  448. package/docs/skills/agent-skills.md +1 -0
  449. package/docs/storage/README.md +8 -6
  450. package/docs/storage/overview.md +24 -19
  451. package/docs/testing/uat/fortemi-live-dataset-uat-plan.md +32 -1
  452. package/docs/verification-contracts.md +15 -0
  453. package/docs/voice/channels.md +59 -0
  454. package/docs/voice/consumers.md +38 -0
  455. package/docs/voice/evaluation-protocol.md +78 -0
  456. package/docs/voice/evidence/channel-literals-development-2026-09-07.json +30 -0
  457. package/docs/voice/evidence/channel-literals-development-2026-09-07.md +25 -0
  458. package/docs/voice/evidence/channel-model-comparison-2026-09-07.json +88 -0
  459. package/docs/voice/evidence/channel-model-comparison-2026-09-07.md +29 -0
  460. package/docs/voice/evidence/core-guidance-development-2026-09-07.json +32 -0
  461. package/docs/voice/evidence/core-guidance-development-2026-09-07.md +23 -0
  462. package/docs/voice/evidence/current-model-expression-2026-09-07.json +99 -0
  463. package/docs/voice/evidence/current-model-expression-2026-09-07.md +26 -0
  464. package/docs/voice/evidence/expression-edits-development-2026-09-07.json +52 -0
  465. package/docs/voice/evidence/expression-edits-development-2026-09-07.md +29 -0
  466. package/docs/voice/evidence/frontier-feedback-ablation-2026-09-07.json +91 -0
  467. package/docs/voice/evidence/frontier-feedback-ablation-2026-09-07.md +34 -0
  468. package/docs/voice/evidence/gemma-native-schema-2026-09-07.json +19 -0
  469. package/docs/voice/evidence/gemma-native-schema-2026-09-07.md +23 -0
  470. package/docs/voice/evidence/harness-correction-2026-09-07.json +240 -0
  471. package/docs/voice/evidence/harness-correction-2026-09-07.md +28 -0
  472. package/docs/voice/evidence/harness-models-2026-09-07.json +37 -0
  473. package/docs/voice/evidence/harness-models-2026-09-07.md +30 -0
  474. package/docs/voice/evidence/model-eligibility.md +7 -0
  475. package/docs/voice/evidence/selector-development-2026-09-07.json +12764 -0
  476. package/docs/voice/evidence/selector-development-2026-09-07.md +32 -0
  477. package/docs/voice/evidence/span-repair-development-2026-09-07.json +27 -0
  478. package/docs/voice/evidence/span-repair-development-2026-09-07.md +11 -0
  479. package/docs/voice/evidence/tinystyler-development-2026-09-07.json +65 -0
  480. package/docs/voice/evidence/tinystyler-development-2026-09-07.md +15 -0
  481. package/docs/voice/exemplar-selection.md +91 -0
  482. package/docs/voice/fidelity.md +23 -0
  483. package/docs/voice/flows/README.md +11 -0
  484. package/docs/voice/flows/voice-critique-correction.flow.yaml +344 -0
  485. package/docs/voice/qualification.md +50 -0
  486. package/docs/voice/receipts-and-migration.md +19 -0
  487. package/docs/voice/revision.md +57 -0
  488. package/docs/voice/writer-profiles.md +63 -0
  489. package/docs/voice/writing-briefs.md +78 -0
  490. package/docs/welcome.html +49 -158
  491. package/docs/yaml-metalanguage.md +1 -1
  492. package/package.json +24 -3
  493. package/prebuilt/fortemi-core/framework/aiwg-fortemi-index-v2.json +1 -1
  494. package/prebuilt/fortemi-core/framework/manifest.json +2 -2
  495. package/schemas/catalog/catalog.json +2 -1
  496. package/schemas/catalog/domains/dataset.json +101 -0
  497. package/schemas/catalog/domains/network-analysis.json +121 -0
  498. package/schemas/catalog/domains/repository-json-schemas.json +20 -0
  499. package/schemas/dataset/fortemi-live-qualification-receipt.v2.schema.json +196 -0
  500. package/schemas/dataset/fortemi-run-receipt/validation-1.0.1/authority.json +12 -0
  501. package/schemas/dataset/fortemi-run-receipt/validation-1.0.1/run-receipt.schema.json +819 -0
  502. package/schemas/mission-protocol/inventory-v1.json +10 -10
  503. package/schemas/network-analysis/analysis-recipe.v1.schema.json +279 -0
  504. package/schemas/network-analysis/governance-record.v1.schema.json +180 -0
  505. package/schemas/network-analysis/packet-evidence.v1.schema.json +487 -0
  506. package/test/fixtures/network-analysis/conformance-report.v1.json +62 -0
  507. package/test/fixtures/network-analysis/manifest.v1.json +68 -0
  508. package/tools/agents/deploy-agents.mjs +7 -3
  509. package/tools/agents/providers/antigravity.mjs +1 -1
  510. package/tools/agents/providers/base.mjs +3 -2
  511. package/tools/agents/providers/deepseek-harness.mjs +66 -0
  512. package/tools/agents/providers/hermes.mjs +1 -1
  513. package/tools/agents/providers/openhuman.mjs +2 -2
  514. package/tools/cli/doctor.mjs +1 -1
  515. package/tools/cli/validate-writing.mjs +20 -14
  516. package/tools/cli/wizard.mjs +1 -1
  517. package/tools/manifest/check-discovery-coverage.mjs +2 -2
  518. package/tools/providers/deepseek-harness-live-smoke.mjs +32 -0
  519. package/tools/providers/deepseek-harness-transport.mjs +358 -0
  520. package/tools/qualification/dataset-fortemi-execute.ts +70 -0
  521. package/tools/ralph-external/index.mjs +3 -3
  522. package/tools/ralph-external/lib/deepseek-harness-adapter.mjs +39 -0
  523. package/tools/ralph-external/lib/provider-adapter.mjs +3 -0
  524. package/tools/writing/writing-validator.mjs +34 -35
package/README.md CHANGED
@@ -1,12 +1,19 @@
1
1
  <div align="center">
2
2
 
3
- <a href="https://aiwg.io"><img src="https://aiwg.io/assets/badges/aiwg-hero-dark.png" alt="AIWG — multi-agent AI framework · one source of truth · 14 provider integrations including Google Antigravity CLI" width="680"></a>
3
+ <a href="https://aiwg.io"><img src="docs/.public/aiwg-readme-hero-v2.png" alt="AIWG — multi-agent AI framework, one
4
+ source of truth; network connecting AI tools" width="1000"></a>
4
5
 
5
6
  # AIWG
6
7
 
7
- **Multi-agent AI framework for 14 provider integrations, including Antigravity, Claude Code, Codex, Copilot, Cursor, Pi, and Warp**
8
+ **Reusable project context and specialist workflows for the AI tools you already use.**
8
9
 
9
- 200+ agents, 109+ CLI commands, 400+ deployable agent/skill/command/rule artifacts, 8 core frameworks, 32 addons, and a 40-plugin Claude Code marketplace. SDLC workflows, digital forensics, research management, marketing operations, media curation, ops infrastructure, knowledge base, and fine-tuning dataset curation — all deployable with one command.
10
+ Plan software, coordinate specialist reviews, prepare campaigns, investigate incidents, organize research, curate
11
+ media, and maintain operational knowledge. AIWG combines agents, skills, rules, templates, and workflow utilities
12
+ around these tasks, adapting them to your existing AI provider.
13
+
14
+ Project artifacts carry decisions from one session to the next. Domain frameworks supply the procedures; addons extend
15
+ them with writing profiles, task loops, memory, testing tools, and other capabilities. The sections below show what
16
+ you can do, how the pieces work together, and how to use them.
10
17
 
11
18
  The simplest setup is to paste this into a supported AI provider:
12
19
 
@@ -23,7 +30,7 @@ system with one self-verifying `aiwg use all` command. That command refreshes
23
30
  the indices, regenerates project context, verifies the resulting deployment,
24
31
  and reports whether a provider reload is actually required.
25
32
 
26
- For secure long-running agents, install AIWG Cockpit with a self-hosted Agentic
33
+ For long-running agents that need an isolated executor, optionally install AIWG Cockpit with a self-hosted Agentic
27
34
  Sandbox executor you control and audit:
28
35
 
29
36
  ```text
@@ -69,7 +76,7 @@ See [Web-Backed AIWG Resources](docs/install/web-backed-resources.md) for source
69
76
  selection, exact-version overrides, cache verification, offline use, and the
70
77
  current framework-graph constraints.
71
78
 
72
- Then ask your AI assistant to set up the project for AIWG. The agent-led setup
79
+ For a larger project, ask your AI assistant to establish project policy as well. The agent-led setup
73
80
  conversation should establish remotes, issue storage, delivery behavior,
74
81
  signing policy, and provider choices; the assistant may call `aiwg setup project`
75
82
  as the underlying CLI helper.
@@ -85,7 +92,7 @@ Agents and stewards setting up AIWG end-to-end should use the
85
92
  [![GitHub Stars](https://img.shields.io/github/stars/jmagly/aiwg?style=flat-square)](https://github.com/jmagly/aiwg/stargazers)
86
93
  [![Node Version](https://img.shields.io/badge/node-%E2%89%A520.0.0-brightgreen?style=flat-square&logo=node.js)](https://nodejs.org)
87
94
  [![TypeScript](https://img.shields.io/badge/TypeScript-5.x-blue?style=flat-square&logo=typescript)](https://www.typescriptlang.org)
88
- [![14 Providers](https://img.shields.io/badge/Providers-14-purple?style=flat-square)](#-platform-support)
95
+ [![15 Providers](https://img.shields.io/badge/Providers-15-purple?style=flat-square)](#platform-support)
89
96
  [![Listed on mcpservers.org](https://mcpservers.org/badge.svg)](https://mcpservers.org/servers/docs-aiwg-io)
90
97
 
91
98
  [![Built With AIWG](https://aiwg.io/assets/badges/built-with-aiwg-dark.png)](https://aiwg.io/badges)
@@ -125,15 +132,16 @@ same scoped rebuild command.
125
132
 
126
133
  If `npm install -g aiwg` fails with `EACCES` while writing to
127
134
  `/usr/local/lib/node_modules/aiwg`, npm is using a system-owned global install
128
- directory. The recommended Mac path is Node 24 through `nvm`, then:
135
+ directory. The [Node.js setup guide](docs/getting-started/install-node.md) covers the supported runtime and
136
+ version-manager choices. After setting up Node, run:
129
137
 
130
138
  ```bash
131
139
  npm install -g aiwg
132
140
  aiwg --version
133
141
  ```
134
142
 
135
- If Node is already installed and you need a quick recovery, use npm's current
136
- user-owned prefix guidance:
143
+ If Node is already installed and you need a quick recovery, one manual alternative is a user-owned npm prefix. Choose
144
+ this only after checking your existing Node version-manager configuration:
137
145
 
138
146
  ```bash
139
147
  npm config set prefix ~/.local
@@ -158,22 +166,33 @@ echo 'export PATH="$(npm config get prefix)/bin:$PATH"' >> ~/.zshrc # or ~/.ba
158
166
  source ~/.zshrc # or restart your shell
159
167
  ```
160
168
 
161
- You can also invoke AIWG without adjusting `PATH` by using `npx aiwg <command>`. For a broader health check — version, deployed providers, missing dependencies, kernel-skill probes — run `aiwg doctor`, which surfaces the same PATH guidance on every invocation if the binary isn't reachable.
169
+ You can also invoke AIWG without adjusting `PATH` by using `npx aiwg <command>`. For a broader health check — version,
170
+ deployed providers, missing dependencies, kernel-skill probes — run `aiwg doctor`. See the [troubleshooting
171
+ guide](docs/troubleshooting/index.md) for the full recovery paths.
162
172
 
163
173
  ---
164
174
 
165
175
  ## What AIWG Is
166
176
 
167
- AIWG is a deployment tool and support utility for AI context. At its core, `aiwg use` copies markdown and YAML source files into the paths each provider reads, so one source of truth works across 14 named provider integrations. A fifteenth `generic` adapter emits portable files for unrecognized or custom harnesses and is not counted as a named integration.
177
+ AIWG gives your AI assistant reusable project context and specialist workflows. Its deployment layer connects those
178
+ instructions to your provider: `aiwg use` copies markdown and YAML source files into the paths each provider reads, so
179
+ one source of truth works across 15 named provider integrations. A sixteenth `generic` adapter emits portable files
180
+ for unrecognized or custom harnesses and is not counted as a named integration.
168
181
 
169
- Around that core, AIWG ships agent-facing utilities for things the base platforms do not handle on their own: persistent artifact memory (`.aiwg/`), background orchestration, autonomous loops, artifact indexing, cost telemetry, health diagnostics, and more. These are tools the agent calls when you ask for something AIWG-shaped — you stay in chat. Most are opt-in. The deployment layer works standalone as plain text files the platform reads natively.
182
+ Around that core, AIWG ships agent-facing utilities for work that benefits from additional structure: persistent
183
+ artifact memory (`.aiwg/`), background orchestration, autonomous loops, artifact indexing, cost telemetry, health
184
+ diagnostics, and more. These are tools the agent calls when the task calls for them — you stay in chat. Most are
185
+ opt-in. The deployment layer works standalone as plain text files the platform reads natively.
170
186
 
171
187
  ### Project scope (recommended) vs user scope (global)
172
188
 
173
189
  `aiwg use` supports project deployments, additive user mirrors, and a
174
190
  user-global bootstrap:
175
191
 
176
- - **Project scope** — default. Run `aiwg use all --provider <provider>` from a project root and the artifacts land in that provider's project paths. One project's agent set never bleeds into another's session. **This is the recommended default for new users.**
192
+ - **Project scope** — default. Run `aiwg use all --provider <provider>` from a project root and the artifacts land in
193
+ that provider's project paths. This keeps project-specific instructions associated with the intended repository.
194
+ Some providers also use user-level surfaces; check the reported deployment scope. **This is the recommended default
195
+ for new users.**
177
196
  - **User scope (additive mirror)** — `aiwg use all --provider <provider> --scope user` keeps the
178
197
  project deployment and mirrors it to `~/.claude/agents/`,
179
198
  `~/.claude/skills/`, etc.
@@ -183,42 +202,59 @@ user-global bootstrap:
183
202
  Use `aiwg regenerate --provider <name>` to wire additional projects without
184
203
  deploying their own skill copies.
185
204
 
186
- The trade-off is real: when the same agent set loads into every session, context from one project can bleed into reasoning about another. Research (REF-720, *Lost in Multi-Turn Conversation*, MSR/Salesforce 2025) measured a 39% capability drop when this happens. The non-blocking project-isolation warning surfaces the trade-off at deploy time so the scope choice is informed. Neither scope is wrong; pick the one that fits the workflow.
205
+ Shared user-level instructions can be useful for personal conventions, while project-local instructions keep a team's
206
+ requirements and decisions with its repository. Review both scopes when a provider uses them together. The installer
207
+ and provider inventory describe where files will go, so you can distinguish instructions that follow you across
208
+ projects from instructions intended for this workspace.
187
209
 
188
210
  See the [Agentic Install Runbook](docs/agentic-install-runbook.md) for the
189
- zero-to-running setup path, and `https://github.com/jmagly/aiwg/blob/main/docs/cli/reference.md` (under `aiwg use` →
211
+ zero-to-running setup path, and the [CLI reference](docs/cli/reference.md) (under `aiwg use` →
190
212
  "Scope models") for the per-provider details and the global-install rough-edge
191
213
  inventory.
192
214
 
193
215
  ## Simple Building Blocks
194
216
 
195
- AIWG ships five primitive artifact types. All are plain text:
217
+ AIWG's workflow source is readable and editable. The main building blocks are:
196
218
 
197
- - **Agents** — specialized personas (Security Auditor, Test Architect) with a scoped toolset
198
- - **Skills** — natural-language workflows the platform auto-invokes on trigger phrases
199
- - **Commands** — explicit slash invocations (`/flow-security-review-cycle`)
200
- - **Rules** — enforcement directives the platform loads into every session
201
- - **Behaviors** — lifecycle hooks that fire on events (pre-write, post-session)
219
+ - **Agents** — specialist role instructions, such as Security Auditor or Test Architect, with defined responsibilities
220
+ and supported tool access.
221
+ - **Skills** — reusable procedures an agent can find from a goal and follow during a task.
222
+ - **Commands** — explicit ways to request a workflow through the provider or CLI.
223
+ - **Rules** — constraints for the assistant to follow, with tool-based checks where configured.
224
+ - **Behaviors** — lifecycle actions and hooks on providers that support them.
225
+ - **Templates** — structures for requirements, briefs, review reports, runbooks, and other outputs.
202
226
 
203
- Each is a single `.md` file with YAML frontmatter. Nothing executes until an AI platform reads it.
227
+ Many assets use Markdown with YAML metadata; hooks and utilities may also include executable scripts or structured
228
+ configuration. The provider determines how each asset is loaded or invoked. A role definition is not a separate model,
229
+ and a written rule is not proof that its constraint was enforced.
204
230
 
205
231
  ## Why It Compounds
206
232
 
207
- Because the primitives are text, they compose without runtime coordination:
233
+ The building blocks become more useful when workflows share their outputs:
234
+
235
+ - A requirements analyst records acceptance criteria that the test engineer can later use to review coverage.
236
+ - A security reviewer reads the same design decision as the implementation agent and records concerns against that decision.
237
+ - A campaign brief gives content writers a shared audience, message, and review criteria.
238
+ - A research note connects a source to the claim it supports, so a later synthesis can revisit the original evidence.
239
+ - A runbook carries verification and recovery steps from planning into an operational change.
208
240
 
209
- - One agent file becomes one member of a **180-agent SDLC team** that reviews architecture, tests, security, and compliance in parallel.
210
- - One skill becomes a **natural-language entry point** — "run security review" routes to the right multi-agent flow on every platform that supports skills.
211
- - One **framework** (SDLC, forensics, marketing) bundles dozens of agents + skills + rules + templates that cross-reference each other. Deploying a framework deploys a working multi-agent ecosystem.
212
- - The `.aiwg/` directory gives those agents a **shared memory** — artifacts from Monday's requirements session are read by Thursday's test design.
213
- - Flows orchestrate **Primary Author → Parallel Reviewers → Synthesizer → Archive** patterns that no single-prompt workflow can match.
241
+ Frameworks package these relationships: agents, skills, rules, and templates reference one another, while `.aiwg/`
242
+ holds the project-specific work they produce. A review can follow **Primary Author → Reviewers → Synthesizer →
243
+ Approval → Archive**, using parallel execution where the provider and task support it.
214
244
 
215
- The leverage is not in any one file. It is that hundreds of small files — each independently readable and editable — snap together into workflows that would otherwise take a bespoke agent platform to build.
245
+ For example, Monday's architecture review can become Thursday's implementation checklist. The second session needs to
246
+ read the saved artifact and check that it is still applicable, but the team has a concrete record to work from rather
247
+ than reconstructing the decision from conversation fragments.
216
248
 
217
- This is also where the research background lives. AIWG implements patterns from cognitive science (Miller 1956, Sweller 1988), multi-agent systems (Jacobs et al. 1991, MetaGPT, AutoGen), and software engineering (Cooper's stage-gate, FAIR Principles, W3C PROV) — applied as file conventions and deployment rules, not as a runtime you depend on.
249
+ Research on structured artifacts, multi-agent review, and recovery informs this design. The [research
250
+ foundations](#research-foundations) section preserves that background separately from claims about AIWG's own
251
+ performance.
218
252
 
219
253
  ## How You Actually Use AIWG
220
254
 
221
- The user surface is the conversation with your AI tool. You install AIWG, deploy a framework, and then talk to the agent normally — "help me start a project", "run a security review", "find me a deploy workflow." The agent does the AIWG-specific work for you.
255
+ The user surface is the conversation with your AI tool. You install AIWG, deploy a framework, and then talk to the
256
+ agent normally — "help me start a project", "run a security review", "find me a deploy workflow." The agent can
257
+ discover the appropriate procedure, perform the task, and report what it verified.
222
258
 
223
259
  The CLI exists mostly for the agent to call under the hood. The commands a user typically runs by hand are a short list:
224
260
 
@@ -230,53 +266,86 @@ The CLI exists mostly for the agent to call under the hood. The commands a user
230
266
  - `aiwg doctor` — health check
231
267
  - `aiwg refresh` — keep the install current
232
268
 
233
- Everything else is agent territory. Discovery (`aiwg discover`), artifact lookup (`aiwg show`), the index, agent loops, mission control, MCP — those are tools the agent invokes during a chat when you ask for something AIWG-shaped. You stay in the conversation; the agent handles the lookups, runs the loops, and reports back.
269
+ Advanced operators can also use the CLI directly; most everyday work can stay in chat. Discovery (`aiwg discover`),
270
+ artifact lookup (`aiwg show`), the index, agent loops, mission control, MCP — those are tools the agent invokes during
271
+ a chat when the task calls for them. The agent handles the lookups, executes the selected workflow, and reports its
272
+ result and limitations.
234
273
 
235
- Turn the agent-side tooling on (it's on by default once you `aiwg use`) when you want persistence, parallelism, or automation. Turn it off and the deployed agents, skills, and rules still work — they are still text files the platform reads natively.
274
+ Deployment connects the supported workflow surface. Optional servers, storage services, and automation paths have
275
+ their own prerequisites; installing assets does not mean all those services are running. You can start with one review
276
+ or document task, then enable additional utilities as the work requires.
236
277
 
237
278
  ## What AIWG Is Not
238
279
 
239
- - **Not a prompt library.** Prompts are the artifacts, not the product. The product is placing the right prompts where the platform finds them.
240
- - **Not an LLM runtime.** AIWG never calls a model. The AI platform you already use does that; AIWG configures what it sees.
241
- - **Not a framework you import into your app.** Nothing is imported at build time. Your project gets a `.aiwg/` directory (artifacts) and a few provider-specific context dirs (deployed copies). Delete them and your app is unchanged.
280
+ AIWG adds context, procedures, and support utilities around the AI tools you use. The deployment core writes
281
+ provider-readable assets; optional orchestration and integration components can invoke provider runtimes or other
282
+ configured services. The [architecture overview](docs/architecture-overview.md) distinguishes those components.
283
+
284
+ It does not replace your provider subscription, your application's runtime, or the review needed before using
285
+ generated work. Most workflows operate on project artifacts and source files rather than requiring your application to
286
+ import an AIWG library. Keep generated configuration and project artifacts under the same review discipline as other
287
+ repository changes.
242
288
 
243
289
  ## Who It's For
244
290
 
245
- If you have used AI coding assistants and thought "this is amazing for small tasks but falls apart on anything complex," AIWG is the missing infrastructure layer that scales AI assistance to multi-week projects.
291
+ AIWG is useful to individual developers, engineering teams, technical leaders, researchers, marketers, and operators
292
+ whose work spans several tasks or sessions. It helps when you need reusable instructions, a shared record of
293
+ decisions, or reviews from more than one perspective.
294
+
295
+ You can start at several levels: a small documentation review, a focused code audit, a campaign brief, an
296
+ investigation plan, or a complete development lifecycle. The [capability guide](docs/overview/capabilities.md) offers
297
+ task-based routes, while this README keeps the broader feature and workflow detail available below.
246
298
 
247
299
  ---
248
300
 
249
301
  ## What Problems Does AIWG Solve?
250
302
 
251
- Base AI assistants (Claude, GPT-4, Copilot without frameworks) have three fundamental limitations:
303
+ AI-assisted projects often need a deliberate way to carry context forward, recover from failed attempts, and make
304
+ review criteria explicit. AIWG provides workflows and artifacts for each of those needs.
252
305
 
253
- ### 1. No Memory Across Sessions
306
+ ### 1. Maintaining Context Across Sessions
254
307
 
255
- Each conversation starts fresh. The assistant has no idea what happened yesterday, what requirements you documented, or what decisions you made last week. You re-explain context every morning.
308
+ Useful decisions can become scattered between conversations, issues, and source files. AIWG workflows save project
309
+ outputs in `.aiwg/`, including requirements, architecture decisions, risk notes, test strategies, and campaign
310
+ material.
256
311
 
257
- **Without AIWG**: Projects stall as context rebuilding eats time. A three-month project requires continuity, not fresh starts every session.
312
+ A later task can consult the relevant artifact and link its own work back to it. The requirements analyst writes a use
313
+ case; the test engineer reads it to identify missing coverage; the implementation review checks whether the behavior
314
+ matches the acceptance criteria. A changed decision can be recorded and propagated through those relationships.
258
315
 
259
- **With AIWG**: The `.aiwg/` directory maintains 50-100+ interconnected artifacts across days, weeks, and months. Later phases build on earlier ones automatically because memory persists. Agents read prior work via `@-mentions` instead of regenerating from scratch.
316
+ The structure also helps an agent select context for a large project. Instead of treating every file as equally
317
+ relevant, a task can begin with a requirement, design record, or source note and follow its supporting references.
318
+ Artifact lookup and indexing utilities help locate those records as the collection grows.
260
319
 
261
- The segmented structure also makes large projects tractable. As code files grow, the project doesn't become harder to reason about — agents load only the slice of memory relevant to the current task (`@requirements/UC-001.md`, `@architecture/sad.md`, `@testing/test-plan.md`) rather than the entire codebase. Each subdirectory is a focused knowledge domain that fits comfortably in context, while cross-references keep everything connected.
320
+ The benefit depends on keeping artifacts current and actually consulting them. A saved file is a reusable source of
321
+ context, not a guarantee that every later response will use it correctly.
262
322
 
263
- The artifact index (`aiwg index`) takes this further. Without any tooling, agents often need to browse 3-6 documents before finding what they need. AIWG's structured artifacts reduce this to 2-3. With the index enabled, agents resolve artifact lookups in one query more often than not — a direct hit on the right requirement, architecture decision, or test case without browsing.
323
+ ### 2. Recovering from Failed Attempts
264
324
 
265
- ### 2. No Recovery Patterns
325
+ A failing test or incomplete task needs a diagnosis, not just another attempt with the same assumptions. Agent loops
326
+ support an execute-and-verify cycle that records failure information, adapts the next attempt, and stops at configured
327
+ limits or escalation conditions.
266
328
 
267
- When AI generates broken code or flawed designs, you manually intervene, explain the problem, and hope the next attempt works. There is no systematic learning from failures, no structured retry, no checkpoint-and-resume.
329
+ Use a loop for a bounded change with an observable completion criterion: fixing a regression, bringing a module under
330
+ test, or carrying out a migration plan. External loop tooling adds process and session recovery where supported. Its
331
+ usefulness depends on the provider, environment, task boundaries, and verification command; unattended execution is
332
+ not a guarantee of completion.
268
333
 
269
- **Without AIWG**: Research shows 47% of AI workflows produce inconsistent outputs without reproducibility constraints (R-LAM, Sureshkumar et al. 2026). Debugging is trial-and-error.
334
+ The record of attempts can help a later session understand what was tried and why it failed. That makes the recovery
335
+ process reviewable even when the agent needs human input or cannot complete the task.
270
336
 
271
- **With AIWG**: The agent loop implements closed-loop self-correction — execute, verify, learn from failure, adapt strategy, retry. External Ralph survives crashes and runs for 6-8+ hours autonomously. Debug memory accumulates failure patterns so the agent doesn't repeat mistakes.
337
+ ### 3. Making Quality Criteria Explicit
272
338
 
273
- ### 3. No Quality Gates
339
+ Different reviews ask different questions. A security review examines exposure and trust boundaries; a performance
340
+ review examines expected load and bottlenecks; a test review checks whether acceptance criteria are exercised; a
341
+ writing review checks audience, clarity, and support for claims.
274
342
 
275
- Base assistants optimize for "sounds plausible" not "actually works." A general assistant critiques security, performance, and maintainability simultaneously — poorly. No domain specialization, no multi-perspective review, no human approval checkpoints.
343
+ AIWG supplies specialist roles and workflows that separate these concerns, combine their findings, and record open
344
+ decisions. Phase gates can check whether the required artifacts and reviews are ready before the work advances.
345
+ Project policy determines who must approve a decision and which checks are required.
276
346
 
277
- **Without AIWG**: Production code ships without architectural review, security validation, or operational feasibility assessment.
278
-
279
- **With AIWG**: 162 specialized agents provide domain expertise — Security Auditor reviews security, Test Architect reviews testability, Performance Engineer reviews scalability. Multi-agent review panels with synthesis. Human-in-the-loop gates at every phase transition. Research shows 84% cost reduction keeping humans on high-stakes decisions versus fully autonomous systems (Agent Laboratory, Schmidgall et al. 2025).
347
+ Multiple reviewers can still share an error. The practical benefit is a clearer review procedure and a saved account
348
+ of the findings, with tests or other independent checks where available.
280
349
 
281
350
  ---
282
351
 
@@ -284,13 +353,17 @@ Base assistants optimize for "sounds plausible" not "actually works." A general
284
353
 
285
354
  ### 1. Memory — Structured Semantic Memory
286
355
 
287
- The `.aiwg/` directory is a persistent artifact repository storing requirements, architecture decisions, test strategies, risk registers, and deployment plans across sessions. This implements Retrieval-Augmented Generation patterns (Lewis et al., 2020) — agents retrieve from an evolving knowledge base rather than regenerating from scratch.
356
+ The `.aiwg/` directory is a persistent artifact repository storing requirements, architecture decisions, test
357
+ strategies, risk registers, and deployment plans across sessions. Artifacts provide retrievable project knowledge that
358
+ can ground later work in recorded decisions and sources.
288
359
 
289
- Each artifact is discoverable via `@-mentions` (e.g., `@.aiwg/requirements/UC-001-login.md`). Context sharing between agents happens through artifacts: the requirements analyst writes use cases, the architecture designer reads them.
360
+ Artifacts can be referenced with `@-mentions` (e.g., `@.aiwg/requirements/UC-001-login.md`). Context sharing between
361
+ agents happens through artifacts: the requirements analyst writes use cases, the architecture designer reads them.
290
362
 
291
363
  ### 2. Reasoning — Multi-Agent Deliberation with Synthesis
292
364
 
293
- Instead of a single general-purpose assistant, AIWG provides 162 specialized agents organized by domain. Complex artifacts go through multi-agent review panels:
365
+ AIWG provides specialist role definitions organized by domain. A workflow can route a complex artifact through the
366
+ reviewers the task requires:
294
367
 
295
368
  ```
296
369
  Architecture Document Creation:
@@ -304,11 +377,14 @@ Architecture Document Creation:
304
377
  4. Human approval gate → accept, iterate, or escalate
305
378
  ```
306
379
 
307
- Research shows 17.9% accuracy improvement with multi-path review on complex tasks (Wang et al., GSM8K benchmarks, 2023). Agent specialization means security review is done by a security specialist, not a generalist.
380
+ Each reviewer receives a defined responsibility and relevant context. The synthesis should resolve duplicate or
381
+ conflicting findings, preserve uncertainty, and identify which conclusions were checked against sources or tests.
382
+ Parallel reviews require the corresponding provider capability and available task budget.
308
383
 
309
384
  ### 3. Learning — Closed-Loop Self-Correction (Ralph)
310
385
 
311
- Ralph executes tasks iteratively, learns from failures, and adapts strategy based on error patterns. Research from Roig (2025) shows recovery capability — not initial correctness — predicts agentic task success.
386
+ Ralph executes tasks iteratively and uses verification results to guide the next attempt. Its task record can preserve
387
+ failure analysis and revised strategies for subsequent iterations.
312
388
 
313
389
  ```
314
390
  Ralph Iteration:
@@ -316,14 +392,16 @@ Ralph Iteration:
316
392
  2. Verify results (tests pass, lint clean, types check)
317
393
  3. If failure: analyze root cause → extract structured learning → adapt strategy
318
394
  4. Log iteration state (checkpoint for resume)
319
- 5. Repeat until success or escalate to human after 3 failed attempts
395
+ 5. Repeat within configured limits; stop or escalate when required
320
396
  ```
321
397
 
322
- External Ralph adds crash resilience: PID file tracking, automatic restart, cross-session persistence. Tasks run for 6-8+ hours surviving terminal disconnects and system reboots.
398
+ External Ralph adds process tracking, session persistence, and recovery controls. For long-running work, define a time
399
+ or iteration budget and inspect the provider-specific recovery behavior; surviving a particular failure depends on how
400
+ the runner and host are configured.
323
401
 
324
402
  ### 4. Verification — Bidirectional Traceability
325
403
 
326
- AIWG maintains links between documentation and code to ensure artifacts stay synchronized:
404
+ AIWG supports links between documentation and code so reviewers can inspect relationships and find drift:
327
405
 
328
406
  ```typescript
329
407
  // src/auth/login.ts
@@ -335,7 +413,9 @@ AIWG maintains links between documentation and code to ensure artifacts stay syn
335
413
  export function authenticateUser(credentials: Credentials): Promise<AuthResult> {
336
414
  ```
337
415
 
338
- Verification types: Doc → Code, Code → Doc, Code → Tests, Citations → Sources. The retrieval-first citation architecture reduces citation hallucination from 56% to 0% (LitLLM benchmarks, ServiceNow 2025).
416
+ Verification can follow Doc → Code, Code → Doc, Code → Tests, and Citations → Sources. These relationships make claims
417
+ easier to inspect; they do not prove that implementation or citations are correct. Ask the workflow to report missing
418
+ targets, inconsistent behavior, and unsupported source claims.
339
419
 
340
420
  ### 5. Planning — Phase Gates with Cognitive Load Management
341
421
 
@@ -346,16 +426,16 @@ Inception → Elaboration → Construction → Transition → Production
346
426
  LOM ABM IOC PR
347
427
  ```
348
428
 
349
- Cognitive load optimization follows Miller's 7±2 limits (1956) and Sweller's worked examples approach (1988):
350
-
351
- - 4 phases (not 12)
352
- - 3-5 artifacts per phase (not 20)
353
- - 5-7 section headings per template (not 15)
354
- - 3-5 reviewers per panel (not 10)
429
+ Phase structure gives a team bounded decisions and review points: establish goals during Inception, evaluate design
430
+ and risks during Elaboration, implement and verify during Construction, and prepare operational handoff during
431
+ Transition. Templates help make the expected outputs visible. The project determines which artifacts, reviewers, and
432
+ approval gates are appropriate; a small task need not use the full lifecycle.
355
433
 
356
434
  ### 6. Style — Controllable Voice Generation
357
435
 
358
- Voice profiles provide continuous control over AI writing style using 12 parameters (formality, technical depth, sentence variety, jargon density, personal tone, humor, directness, examples ratio, uncertainty acknowledgment, opinion strength, transition style, authenticity markers).
436
+ Voice profiles describe writing preferences such as formality, technical depth, directness, examples, uncertainty, and
437
+ sentence variation. They give the assistant a reusable style specification that can be reviewed against the audience
438
+ and document purpose.
359
439
 
360
440
  Built-in voices: `technical-authority` (docs, RFCs), `friendly-explainer` (tutorials), `executive-brief` (summaries), `casual-conversational` (blogs, social). Create custom voices from your existing content with `/voice-create`.
361
441
 
@@ -363,7 +443,13 @@ Built-in voices: `technical-authority` (docs, RFCs), `friendly-explainer` (tutor
363
443
 
364
444
  ## A Real Project Walkthrough
365
445
 
366
- Here is how the six components work together across a project lifecycle. How long each phase takes depends entirely on the project — AIWG is a force multiplier, not a clock. Most projects arrive at a complete, reviewed document set in hours to a day. What takes time is the human work that matters: reviewing, editing, and making decisions. The more input your team provides, the better the output. AIWG memory lets operators participate through the tools they already use — industry-standard documents and templates, issues, and knowledge bases.
446
+ The following illustrative customer-portal project shows how the components connect across a lifecycle. The commands
447
+ are provider-facing workflow examples, not a report of a measured project outcome. Natural-language requests can
448
+ select the same procedures when the provider does not expose the shown slash syntax.
449
+
450
+ Start with a bounded goal, agree on acceptance criteria, and inspect the artifacts at each step. The time and review
451
+ effort depend on the project, source quality, tools, and decisions involved. Smaller changes can enter at the phase
452
+ that fits their current state rather than repeating the entire lifecycle.
367
453
 
368
454
  ### Inception
369
455
 
@@ -409,25 +495,31 @@ Here is how the six components work together across a project lifecycle. How lon
409
495
  ```
410
496
 
411
497
  **Planning**: Deployment checklist — monitoring, rollback plan, incident response
412
- **Learning**: Ralph retries deployment steps if validation fails
498
+ **Learning**: Failed validation produces a diagnosis and a revised plan; retries follow the operation's recovery and
499
+ approval requirements
413
500
  **Verification**: Deployment scripts reference architecture (which services, what order)
414
501
  **Human Gate**: Operations team reviews deployment plan → approves production release
415
502
 
416
503
  ---
417
504
 
418
- ## Quantified Claims and Evidence
505
+ ## Claims, Evaluation, and Evidence
419
506
 
420
- AIWG makes specific, falsifiable claims backed by peer-reviewed research:
507
+ Evaluate a workflow against the task it is meant to support. AIWG provides structures for review, traceability, and
508
+ recovery; it does not promise a fixed cost saving, perfect citations, or error-free execution.
421
509
 
422
- | Claim | Evidence | Source |
423
- |-------|----------|--------|
424
- | 84% cost reduction with human-in-the-loop vs fully autonomous | Agent Laboratory study | Schmidgall et al. (2025) |
425
- | 47% workflow failure rate without reproducibility constraints | R-LAM evaluation | Sureshkumar et al. (2026) |
426
- | 0% citation hallucination with retrieval-first vs 56% generation-only | LitLLM benchmarks | ServiceNow (2025) |
427
- | 17.9% accuracy improvement with multi-path review | GSM8K benchmarks | Wang et al. (2023) |
428
- | 18.5x improvement with tree search on planning tasks | Game of 24 results | Yao et al. (2023) |
510
+ | Capability to evaluate | Evidence to collect | Useful comparison |
511
+ |------------------------|---------------------|-------------------|
512
+ | Persistent project context | Whether a later session reads and applies a prior decision | The same kind of task with the team's usual handoff |
513
+ | Specialist review | Correct findings, missed issues, and reviewer effort | A comparable review using the existing process |
514
+ | Task recovery | Failure diagnosis, attempts, limits, and final verification | Similar failures handled without the loop |
515
+ | Citation checking | Source existence and support for each material claim | Manual source inspection |
516
+ | Structured planning | Completeness and usefulness of the resulting plan | The team's existing planning artifact |
517
+ | Cost and throughput | Model calls, elapsed time, and human review effort | Comparable tasks with the same acceptance criteria |
429
518
 
430
- Full references: [docs/research/](docs/research/)
519
+ Research results from Agent Laboratory, self-consistency, tree search, and other systems inform the design. Their
520
+ percentages and benchmark scores are not measurements of AIWG. The [research foundations](#research-foundations) and
521
+ [reading list](docs/overview/reading-list.md) retain the underlying sources. The [executive
522
+ brief](docs/overview/executive-brief.md) describes a practical pilot.
431
523
 
432
524
  ---
433
525
 
@@ -441,13 +533,17 @@ Multi-week or multi-month projects where requirements evolve, multiple stakehold
441
533
 
442
534
  ### Not the Best Fit
443
535
 
444
- Single-session tasks where no memory is needed, quality gates are overkill, and overhead exceeds value.
536
+ A full lifecycle is usually unnecessary for a one-off question that needs no shared context, saved artifact, or
537
+ follow-up. A focused writing, lookup, or review capability can still be useful without introducing every phase and
538
+ gate.
445
539
 
446
540
  **Examples**: "Write a Python script to parse this CSV," "Fix this typo," "Explain how this code works."
447
541
 
448
542
  ### The Trade-off
449
543
 
450
- AIWG adds structure (templates, phases, gates) that slows trivial tasks but scales to complex multi-week workflows. If your project fits in a single conversation, use a base assistant. If it spans days, weeks, or months, AIWG provides the infrastructure to maintain quality and context.
544
+ Match the workflow to the task. Use a bounded skill for a small result, a saved artifact when work needs to carry
545
+ forward, and a phase-based process when coordination and review warrant it. Additional context, reviewers, or
546
+ verification steps can add model calls and human effort; judge their value against the outcome you need.
451
547
 
452
548
  ```
453
549
  User intent → AIWG CLI → Deploy agents + rules + templates → AI platform
@@ -457,11 +553,11 @@ User intent → AIWG CLI → Deploy agents + rules + templates → AI platform
457
553
  │ Cursor / Warp / Factory /
458
554
  ▼ OpenCode / Codex / Devin Desktop
459
555
  ┌──────────────┐
460
- │ 188 Agents │ Specialized AI personas with domain expertise
461
- │ 50 Commands │ CLI + slash commands for workflow automation
462
- │ 128 Skills │ Natural language workflow triggers
463
- │ 35 Rules │ Enforcement patterns (security, quality, anti-laziness)
464
- │ 334 Templates│ SDLC artifact templates with progressive disclosure
556
+ │ Agents │ Specialized AI personas with domain expertise
557
+ │ Commands │ CLI + slash commands for workflow automation
558
+ │ Skills │ Natural language workflow triggers
559
+ │ Rules │ Enforcement patterns (security, quality, anti-laziness)
560
+ │ Templates │ SDLC artifact templates with progressive disclosure
465
561
  └──────────────┘
466
562
  │
467
563
  ▼
@@ -474,17 +570,19 @@ User intent → AIWG CLI → Deploy agents + rules + templates → AI platform
474
570
 
475
571
  > For visual diagrams of AIWG's architecture, deploy flow, and discovery model, see [`docs/architecture-overview.md`](docs/architecture-overview.md). The prose walkthrough lives in [`docs/how-it-works.md`](docs/how-it-works.md).
476
572
 
477
- **At a glance** — AIWG is a deploy-time tool. `aiwg use` copies plain-text files into your AI platform's native directories and exits. Nothing runs in the background; the AI platform's own loader handles everything from there.
573
+ **At a glance** — the deployment layer copies instructions into provider-readable locations, connects project context,
574
+ and reports verification. The provider loads those instructions. Optional runtime components, such as artifact
575
+ services or orchestration tools, perform additional work when configured and invoked.
478
576
 
479
577
  ```mermaid
480
578
  flowchart LR
481
579
  subgraph Source["AIWG framework source"]
482
580
  direction TB
483
- KERN[25 kernel skills<br/>within provider listing budgets]
484
- STD[~455 standard skills<br/>read from $AIWG_ROOT]
485
- AGENT[200+ agents]
486
- RULES[60+ rules]
487
- TPL[100+ templates]
581
+ KERN[Kernel skills<br/>within provider listing budgets]
582
+ STD[Standard skills<br/>read from $AIWG_ROOT]
583
+ AGENT[Specialist agents]
584
+ RULES[Workflow rules]
585
+ TPL[Artifact templates]
488
586
  end
489
587
 
490
588
  CLI([aiwg use all<br/>--provider X]) --> DEPLOY
@@ -538,22 +636,39 @@ The orchestration pattern: **Primary Author → Parallel Reviewers → Synthesiz
538
636
 
539
637
  ## Features
540
638
 
541
- - **188 specialized agents** — domain experts across testing, security, architecture, DevOps, cloud, frontend, backend, data engineering, documentation, and more
542
- - **50 CLI commands** — framework deployment, project scaffolding, iterative execution, metrics, reproducibility validation
543
- - **128 workflow skills** — natural language triggers for regression testing, forensics, voice profiles, quality gates, and CI/CD integration
544
- - **35 enforcement rules** — anti-laziness detection, token security, citation integrity, executable feedback, failure mitigation across 6 LLM archetypes
545
- - **334 artifact templates** — progressive disclosure templates for requirements, architecture, testing, security, deployment, and more
546
- - **Multi-provider support** — deploy to Google Antigravity CLI, Claude Code, OpenAI Codex, GitHub Copilot, Cursor, Factory AI, Hermes, OpenCode, OpenClaw, OpenHuman, [Pi Coding Agent](https://pi.dev/), Oh My Pi, Warp Terminal, and Devin Desktop
547
- - **8 core frameworks + training marketplace package** — SDLC, Digital Forensics, Marketing Operations, Research Management, Media Curation, Ops Infrastructure, Knowledge Base, Security Engineering, plus [`aiwg-training`](https://github.com/jmagly/aiwg-training) for fine-tuning dataset curation (corpus-to-dataset pipeline with DPO/KTO/ORPO/SimPO export)
548
- - **32 addons** — compound memory, line memory, llm-wiki (Obsidian-native knowledge base), RLM recursive decomposition, fleet operations, browser control, testing quality, and more
549
- - **40 Claude Code plugins** — the complete framework and addon catalog is installable independently from the AIWG marketplace
550
- - **Agent Loop** — iterative task execution with automatic error recovery and crash resilience (6-8 hour sessions)
551
- - **RLM addon** — recursive context decomposition for processing 10M+ tokens via sub-agent delegation
552
- - **YAML metalanguage** — declarative schema-validated workflow definitions (JSON Schema 2020-12)
553
- - **MCP server** — Model Context Protocol integration for tool-based AI workflows
554
- - **Bidirectional traceability** — @-mention system linking requirements → architecture → code → tests
555
- - **FAIR-aligned artifacts** — W3C PROV provenance, GRADE quality assessment, persistent REF-XXX identifiers
556
- - **Reproducibility validation** — deterministic execution modes, checkpoints, configuration snapshots
639
+ - **Specialist agents** — roles for architecture, implementation, testing, security, cloud, data engineering,
640
+ research, content, and operations.
641
+ - **Workflow skills and commands** — discoverable procedures for reviews, intake, research, curation, planning, and delivery.
642
+ - **Rules and review criteria** — instructions for preserving work, handling sensitive configuration, checking claims,
643
+ and reporting verification.
644
+ - **Artifact templates** — structured requirements, design decisions, campaign briefs, source notes, runbooks, and
645
+ review reports.
646
+ - **Multi-provider deployment** — Google Antigravity CLI, Claude Code, OpenAI Codex, GitHub Copilot, Cursor, DeepSeek
647
+ Harness, Factory AI, Hermes, OpenCode, OpenClaw, OpenHuman, Pi Coding Agent, Oh My Pi, Warp Terminal, and Devin
648
+ Desktop.
649
+ - **Domain frameworks** — software development, forensics, marketing, research, media curation, operations, knowledge
650
+ base, and security engineering.
651
+ - **Dataset workflows** — assessment, indexing, lineage, synchronization, and retirement through dataset intelligence;
652
+ the separate [`aiwg-training`](https://github.com/jmagly/aiwg-training) project covers training-data curation and
653
+ exports.
654
+ - **Memory addons** — compound memory, line memory, wiki-oriented knowledge, and artifact lookup for different
655
+ persistence needs.
656
+ - **Writing and voice tools** — reusable voice profiles, context-sensitive diagnostics, revision workflows, and
657
+ alternatives for content generation.
658
+ - **Testing quality** — test strategy, mutation testing, flaky-test review, and verification procedures.
659
+ - **Agent loops** — bounded execution, failure analysis, checkpoints, and supported process recovery.
660
+ - **RLM** — recursive context decomposition for tasks whose source material needs to be divided into smaller working sets.
661
+ - **YAML metalanguage** — structured workflow and artifact definitions with schema-oriented validation.
662
+ - **MCP integration** — tools and resources exposed through configured servers and provider connections.
663
+ - **Traceability and provenance** — relationships between requirements, code, tests, sources, and generated artifacts.
664
+ - **Session history and diagnostics** — import and inspect prior AI work, check deployment health, and diagnose
665
+ provider wiring.
666
+ - **Project-local extensions and marketplace delivery** — keep custom instructions with the project and package
667
+ reusable capabilities through the appropriate distribution path.
668
+
669
+ The [framework and addon catalog](#what-you-get) below describes these capabilities in more detail. Compatibility and
670
+ execution requirements are explicit in the [provider inventory](docs/providers/provider-inventory.md) and [CLI
671
+ reference](docs/cli/reference.md).
557
672
 
558
673
  ---
559
674
 
@@ -561,53 +676,69 @@ The orchestration pattern: **Primary Author → Parallel Reviewers → Synthesiz
561
676
 
562
677
  > **Prerequisites:** Node.js >=20.0.0 and an AI platform (Claude Code, GitHub Copilot, Cursor, Warp Terminal, or others). New installs should prefer Node 24. See [Prerequisites Guide](docs/getting-started/prerequisites.md) for details.
563
678
 
564
- > **Verifying releases (v2026.5.3+):** Every AIWG release ships with Sigstore-anchored npm provenance, a signed git tag, a cosign keyless tarball signature, and a signed CycloneDX SBOM. Verification is optional but recommended:
565
- >
566
- > ```bash
567
- > npm view aiwg@2026.5.3 --json | jq .dist.attestations
568
- > ```
569
- >
570
- > Full walkthrough at [`docs/releases/verifying.md`](docs/releases/verifying.md). Adopt the same pattern for your own packages: [`docs/security/supply-chain-hardening.md`](docs/security/supply-chain-hardening.md).
679
+ > **Release verification:** Inspect the provenance and signature material for the release you install. The
680
+ [verification guide](docs/releases/verifying.md) describes the available artifacts and commands.
571
681
 
572
682
  ### Install & Deploy
573
683
 
684
+ The prompt-led installer at the top of this README is the canonical beginner path. For manual setup:
685
+
574
686
  ```bash
575
- # Install globally
576
- npm install -g aiwg
687
+ npm i -g aiwg
688
+ cd /path/to/your/project
689
+ aiwg use all --provider claude # replace claude with your provider selector
690
+ ```
577
691
 
578
- # Deploy to your project
579
- cd your-project
580
- aiwg use sdlc # Full SDLC framework (98 agents, 38 rules, 200+ templates)
581
- aiwg use forensics # Digital forensics & incident response (13 agents, 10 skills)
582
- aiwg use marketing # Marketing operations (37 agents, 87+ templates)
583
- aiwg use media-curator # Media archive management (6 agents, 9 commands)
584
- aiwg use research # Research workflow automation (8 agents, 8-stage pipeline)
585
- aiwg use civic-action # Cited civic review and publication-preparation kit
586
- aiwg use rlm # RLM addon (recursive context decomposition)
587
- aiwg use all # Everything
588
-
589
- # Existing projects: preview, then transactionally extract the canonical graph
590
- aiwg regenerate --existing-project --dry-run
591
- aiwg regenerate --existing-project --apply
592
- aiwg workspace-context doctor
692
+ Deployment refreshes the shared context and reports verification and any required reload. Follow that result, then ask
693
+ the agent to check the intended project and its AIWG connection. The [manual installation
694
+ reference](docs/cli/install-and-repair.md) covers the terminal path in detail.
593
695
 
594
- # Recommended default: inspect project state and select the safe branch
696
+ For a deliberately narrower deployment, choose the relevant framework or addon instead of `all`. These are
697
+ alternatives, not a sequence of required setup steps:
698
+
699
+ ```bash
700
+ aiwg use sdlc --provider claude # Software development
701
+ aiwg use forensics --provider claude # Investigation workflows
702
+ aiwg use marketing --provider claude # Campaign and content work
703
+ aiwg use media-curator --provider claude # Media collections
704
+ aiwg use research --provider claude # Research artifacts
705
+ aiwg use civic-action --provider claude # Civic review and preparation
706
+ aiwg use rlm --provider claude # Context decomposition
707
+ ```
708
+
709
+ For maintenance or an existing workspace that needs context migration, preview the relevant regeneration branch rather
710
+ than treating every branch as installation:
711
+
712
+ ```bash
595
713
  aiwg regenerate --dry-run
596
714
  aiwg regenerate
597
715
 
598
- # Fresh or already-migrated projects: ordinary canonical refresh
599
- aiwg regenerate --workspace
716
+ # Existing-project extraction, when that is the intended operation:
717
+ aiwg regenerate --existing-project --dry-run
718
+ aiwg regenerate --existing-project --apply
719
+ aiwg workspace-context doctor
720
+ ```
721
+
722
+ The [regeneration guide](docs/regenerate-guide.md) also covers canonical refresh and legacy compatibility. Use the
723
+ branch that matches the workspace state. To scaffold a new project rather than connect the current one, see the
724
+ [new-project guide](docs/getting-started/new-project.md).
600
725
 
601
- # Legacy compatibility only: inline AIWG context in provider startup files
602
- aiwg regenerate --full-inject
726
+ ### Get a First Useful Result
603
727
 
604
- # Or scaffold a new project
605
- aiwg new my-project
728
+ After setup, ask your agent:
606
729
 
607
- # Check installation health
608
- aiwg doctor
730
+ ```text
731
+ Use AIWG to review this project's README for unclear positioning and missing
732
+ onboarding steps. Save a report at
733
+ .aiwg/marketing/brand/audit/readme-review.md with file references and the
734
+ three highest-priority fixes. Leave the README unchanged.
609
735
  ```
610
736
 
737
+ Open the report and check the source references, reader impact, and proposed fixes. In a later session, ask the agent
738
+ to read that report and implement the first agreed change. The [first-result
739
+ walkthrough](docs/getting-started/just-try-it.md) includes an illustrative finding and alternative tasks. Setup
740
+ readiness is the prerequisite; a useful artifact is what lets you assess the workflow.
741
+
611
742
  ### Customize Without Forking
612
743
 
613
744
  Author project-specific rules, skills, agents, addons, or frameworks
@@ -637,7 +768,7 @@ The bundle is **byte-identical** in shape to its upstream form, so
637
768
  /plugin install compound-memory@aiwg
638
769
  ```
639
770
 
640
- The marketplace contains 40 independently packaged framework and addon
771
+ The marketplace contains independently packaged framework and addon
641
772
  plugins, so you can install only the capabilities a Claude Code workspace
642
773
  needs. Source-distributed opt-in addons such as Civic Action deploy with
643
774
  `aiwg use civic-action` and do not imply a marketplace wrapper.
@@ -657,6 +788,8 @@ aiwg use all --provider devin # Devin Desktop
657
788
  aiwg use all --provider openclaw # OpenClaw
658
789
  aiwg use all --provider hermes # Hermes
659
790
  aiwg use all --provider openhuman # OpenHuman
791
+ aiwg use all --provider pi # Pi Coding Agent
792
+ aiwg use all --provider omp # Oh My Pi
660
793
  ```
661
794
 
662
795
  `all` means the complete deployable end-user surface. It intentionally omits
@@ -665,11 +798,13 @@ directly.
665
798
 
666
799
  ### First-Party Integrators
667
800
 
668
- Some partners ship AIWG bundled in their runtime — no `aiwg use` step required. Install the partner tool and AIWG is already wired up.
801
+ AIWG can also be distributed through another runtime. Check the integrator's version and included assets before
802
+ assuming that its bundled surface matches a standalone AIWG install. Follow that runtime's setup instructions, then
803
+ verify the project connection.
669
804
 
670
805
  | Partner | Install | What you get |
671
806
  |---------|---------|--------------|
672
- | **[Omnius](https://www.npmjs.com/package/omnius)** | `npm i -g omnius` | AIWG framework set, skills, agents, and rules embedded in the Omnius autonomous coding agent. Discoverable through `aiwg discover` from inside Omnius sessions, surfaceable through Omnius's MCP and REST bridges. |
807
+ | **[Omnius](https://www.npmjs.com/package/omnius)** | `npm i -g omnius` | An integration path for AIWG assets in an autonomous coding runtime. Consult the package documentation for the bundled version, supported asset surface, and setup requirements. |
673
808
 
674
809
  If you ship a product that bundles AIWG and want to be listed here, open an issue at https://github.com/jmagly/aiwg/issues.
675
810
 
@@ -677,163 +812,200 @@ If you ship a product that bundles AIWG and want to be listed here, open an issu
677
812
 
678
813
  ## What You Get
679
814
 
680
- ### Frameworks (8)
681
-
682
- | Framework | Agents | Templates | What It Does |
683
- |-----------|--------|-----------|--------------|
684
- | **[SDLC Complete](agentic/code/frameworks/sdlc-complete/)** | 98 | 200+ | Full software development lifecycle — Inception through Production with multi-agent orchestration, quality gates, and DORA metrics |
685
- | **[Forensics Complete](agentic/code/frameworks/forensics-complete/)** | 13 | 8 | Digital forensics and incident response — evidence acquisition, timeline reconstruction, IOC extraction, Sigma rule hunting. NIST SP 800-86, MITRE ATT&CK, STIX 2.1 |
686
- | **[Media/Marketing Kit](agentic/code/frameworks/media-marketing-kit/)** | 37 | 87+ | End-to-end marketing operations — strategy, content creation, campaign management, brand compliance, analytics, and reporting |
687
- | **[Media Curator](agentic/code/frameworks/media-curator/)** | 6 | — | Intelligent media archive management — discography analysis, source discovery, quality filtering, transcript sidecars, research handoff preparation, multi-platform export (Plex, Jellyfin, MPD) |
688
- | **[Research Complete](agentic/code/frameworks/research-complete/)** | 8 | 6 | Academic research automation — paper discovery, citation management, RAG-based summarization, GRADE quality scoring, FAIR compliance, W3C PROV provenance |
689
- | **[Knowledge Base](agentic/code/frameworks/knowledge-base/)** | — | 5 | General-purpose LLM-assisted wiki — source ingest, entity/concept pages, source summaries, comparisons, syntheses, health checks, and emergent taxonomy |
690
- | **[Ops Complete](agentic/code/frameworks/ops-complete/)** | 12 | 3 | Operational infrastructure — incident management, runbooks, troubleshooting workflows |
691
- | **[Security Engineering](agentic/code/frameworks/security-engineering/)** | 2 | 5 | Applied security beyond STRIDE/OWASP — cryptographic primitive selection, chain-of-trust integrity, authentication-factor architecture, degraded-mode design, runtime secret hygiene, supply-chain trust, physical-access threats. Pattern-based, product-agnostic |
692
-
693
- ### Addons (32)
694
-
695
- | Addon | What It Does |
696
- |-------|--------------|
697
- | **[Civic Action](agentic/code/addons/civic-action/)** | Cited, opt-in decision support for source review, records planning, meeting reconciliation, public-technology research, local-resource profiles, corrections, and publication preparation |
698
- | **[RLM](agentic/code/addons/rlm/)** | Recursive context decomposition — process 10M+ tokens via sub-agent delegation with parallel fan-out |
699
- | **[Writing Quality](agentic/code/addons/writing-quality/)** | Content validation, AI pattern detection, authentic voice enforcement |
700
- | **[Testing Quality](agentic/code/addons/testing-quality/)** | TDD enforcement, mutation testing, flaky test detection and repair |
701
- | **[Voice Framework](agentic/code/addons/voice-framework/)** | 4 built-in voice profiles (technical-authority, friendly-explainer, executive-brief, casual-conversational) with create/blend/apply skills |
702
- | **[UAT-MCP Toolkit](agentic/code/addons/uat-mcp/)** | User acceptance testing with MCP-powered test execution, coverage tracking, and regression detection |
703
- | **[AIWG Evals](agentic/code/addons/aiwg-evals/)** | Agent evaluation framework — archetype resistance testing (Roig 2025), performance benchmarks, quality scoring |
704
- | **[Agent Loop](agentic/code/addons/agent-loop/)** | Iterative task execution engine (`aiwg ralph` / `aiwg agent-loop-ext`) — automatic error recovery, crash resilience, completion tracking |
705
- | **[Agentic Installer](agentic/code/addons/agentic-installer/)** | `setup.aiwg.io/v1` SetupManifest installer — cross-platform install workflows with recovery |
706
- | **[AIWG Dev](agentic/code/addons/aiwg-dev/)** | AIWG development tooling — extension scaffolding, local-source dev mode |
707
- | **[Daemon](agentic/code/addons/daemon/)** | Persistent daemon mode — background sessions, task queue, health monitoring |
708
- | **[Compound Memory](agentic/code/addons/compound-memory/)** | Persistent context architecture combining immutable raw inputs, linked wiki knowledge, line-memory facts, generated outputs, and governed maintenance workflows |
709
- | **[Line Memory](agentic/code/addons/line-memory/)** | Durable append-only facts with retrieval, reinforcement, decay, pruning, and concurrent process safety |
710
- | **[LLM Wiki](agentic/code/addons/llm-wiki/)** | Obsidian-native knowledge base for LLM agents — semantic linking, vault integration |
711
- | **[NLP Prod](agentic/code/addons/nlp-prod/)** | Production NLP pipelines — entity extraction, classification, summarization |
712
- | **[Prose Integration](agentic/code/addons/prose-integration/)** | OpenProse contract grammar integration — declarative service contracts |
713
- | **[Semantic Memory](agentic/code/addons/semantic-memory/)** | Semantic memory kernel — query, capture, lifecycle management for agent memory |
714
- | **[Context Curator](agentic/code/addons/context-curator/)** | Context pre-filtering to remove distractors — production-grade agent reliability |
715
- | **[Verbalized Sampling](agentic/code/addons/verbalized-sampling/)** | Probability distribution prompting — 1.6-2.1x output diversity improvement |
716
- | **[Guided Implementation](agentic/code/addons/guided-implementation/)** | Bounded iteration control for issue-to-code automation |
717
- | **[Skill Factory](agentic/code/addons/skill-factory/)** | Dynamic skill generation and packaging at runtime |
718
- | **[Doc Intelligence](agentic/code/addons/doc-intelligence/)** | Document analysis, PDF extraction, documentation site scraping |
719
- | **[Color Palette](agentic/code/addons/color-palette/)** | WCAG-compliant color palette generation with trend research |
720
- | **[Auto Memory](agentic/code/addons/auto-memory/)** | Automatic memory seed templates for new project context initialization |
721
- | **[Agent Persistence](agentic/code/addons/agent-persistence/)** | Agent state management for session continuity |
722
- | **[AIWG Hooks](agentic/code/addons/aiwg-hooks/)** | Lifecycle event handlers — pre-session, post-write, workflow tracing |
723
- | **[AIWG Utils](agentic/code/addons/aiwg-utils/)** | Core meta-utilities (auto-installed with any framework) |
724
- | **[Droid Bridge](agentic/code/addons/droid-bridge/)** | Factory Droid orchestration — multi-platform agent bridge |
725
- | **[AIWG Fleet](agentic/code/addons/aiwg-fleet/)** | Governed multi-project maintenance with repository discovery, policy-aware planning, approval gates, and auditable execution |
726
- | **[Browser Control](agentic/code/addons/browser-control/)** | Permission-aware browser automation for user-controlled sessions through Playwright MCP |
727
- | **[Twelve-Factor](agentic/code/addons/twelve-factor/)** | Evidence-based application architecture review against Twelve-Factor and modern 12+ Factor criteria |
728
- | **[Star Prompt](agentic/code/addons/star-prompt/)** | Repository star prompt for success celebration |
815
+ AIWG installs reusable context, specialist agents, workflow skills, rules, and
816
+ artifact templates into the AI tools your team already uses. This fragment keeps
817
+ the older README's broad inventory shape while updating claims against the
818
+ current repository. Counts shown in framework rows are source-file counts from
819
+ this working tree; addon rows omit totals because several addons expose
820
+ capabilities through manifests, docs, scripts, or nested skill packages.
821
+
822
+ ### Frameworks
823
+
824
+ | Framework | Source Snapshot | What It Helps You Do |
825
+ |-----------|-----------------|----------------------|
826
+ | **[SDLC Complete](agentic/code/frameworks/sdlc-complete/)** | 100 agents, 116 skills, 217 templates, 39 rules, 12 commands, 8 flows | Run a full software delivery lifecycle from intake through transition with phase gates, planning artifacts, implementation support, test strategy, deployment handoff, and maintenance workflows |
827
+ | **[Forensics Complete](agentic/code/frameworks/forensics-complete/)** | 13 agents, 20 skills, 12 templates, 4 rules | Preserve and analyze incident evidence through scoping, triage, acquisition, log review, persistence hunting, timeline building, IOC extraction, and reporting |
828
+ | **[Media/Marketing Kit](agentic/code/frameworks/media-marketing-kit/)** | 38 agents, 34 skills, 97 templates, 2 flows | Plan, produce, review, publish, and analyze marketing campaigns with reusable briefs, brand/legal gates, channel assets, and performance artifacts |
829
+ | **[Media Curator](agentic/code/frameworks/media-curator/)** | 6 agents, 21 skills | Assess mixed media collections, research sources, acquire approved material, tag metadata, verify integrity, create transcript sidecars, and prepare exports or research handoffs |
830
+ | **[Research Complete](agentic/code/frameworks/research-complete/)** | 8 agents, 41 skills, 16 templates | Turn literature searches and PDFs into reviewable research artifacts: source records, grounded summaries, citation work, GRADE/FAIR-style quality checks, gap notes, and provenance |
831
+ | **[Knowledge Base](agentic/code/frameworks/knowledge-base/)** | 3 skills, 5 templates | Build a linked AI-assisted wiki from loose sources, notes, entities, concepts, comparisons, and synthesis pages without forcing formal literature-review overhead |
832
+ | **[Ops Complete](agentic/code/frameworks/ops-complete/)** | 12 agents, 1 skill, 17 templates, 6 rules | Convert operational procedures into executable runbooks, inventories, incident reports, troubleshooting trees, and extension-backed ops workflows |
833
+ | **[Security Engineering](agentic/code/frameworks/security-engineering/)** | 2 agents, 27 skills, 7 templates, 13 rules | Make applied security decisions for crypto primitives, chains of trust, auth factors, degraded modes, runtime secrets, supply-chain trust, physical threats, and DFIR readiness |
834
+ | **[Validation Complete](agentic/code/frameworks/validation-complete/)** | 1 skill | Add focused validation workflow support where a project needs reviewable checks without adopting a full lifecycle framework |
835
+
836
+ Start with [Install, Connect, and Verify](docs/getting-started/install-connect-verify.md),
837
+ then deploy a framework with `aiwg use <framework>`. The [capability reference](docs/cli/reference.md)
838
+ lists the current framework names accepted by the CLI.
839
+
840
+ ### Addons
841
+
842
+ | Addon | What It Helps You Do |
843
+ |-------|----------------------|
844
+ | **[AIWG Utils](agentic/code/addons/aiwg-utils/)** | Shared rules, discovery helpers, regeneration support, mention tooling, workspace maintenance, and stewardship primitives used across AIWG |
845
+ | **[Agent Loop](agentic/code/addons/agent-loop/)** | Run bounded iterative agent loops with recovery, reflection, completion tracking, and CLI surfaces such as `aiwg ralph` |
846
+ | **[RLM](agentic/code/addons/rlm/)** | Decompose large codebases or document corpora into smaller reviewed slices through recursive planning and subtask execution |
847
+ | **[Composition Engine](agentic/code/addons/composition-engine/)** | Define and validate provider-neutral Flow graph contracts for composed workflows |
848
+ | **[Graph Pattern](agentic/code/addons/graph-pattern/)** | Add an optional graph-oriented profile over AIWG Flow for conditional routes, reducers, and graph validation |
849
+ | **[Orchestration Topology Lab](agentic/code/addons/orchestration-topology-lab/)** | Compare single-agent, bounded-parallel, and planner-worker orchestration topologies using local fixtures and explicit evidence |
850
+ | **[Guided Implementation](agentic/code/addons/guided-implementation/)** | Keep issue-to-code work inside a bounded retry loop with validation after each attempt and structured escalation when needed |
851
+ | **[Daemon](agentic/code/addons/daemon/)** | Run opt-in persistent session support for background tasks, queues, health checks, and scheduler integration |
852
+ | **[Agentic Installer](agentic/code/addons/agentic-installer/)** | Use `setup.aiwg.io/v1` SetupManifest files for reproducible, agent-driven install workflows with recovery paths |
853
+ | **[AIWG Dev](agentic/code/addons/aiwg-dev/)** | Scaffold and validate AIWG source packages, skills, agents, commands, and rules; install explicitly for contributor work |
854
+ | **[Skill Factory](agentic/code/addons/skill-factory/)** | Build, enhance, validate, and package skills through a dedicated skill-authoring workflow |
855
+ | **[AIWG Evals](agentic/code/addons/aiwg-evals/)** | Run agent and workflow evaluation patterns with explicit benchmark inputs and quality scoring |
856
+ | **[Monitorability Red Team](agentic/code/addons/monitorability-red-team/)** | Exercise synthetic local fixtures that expose multi-agent monitoring limits and evidence blind spots |
857
+ | **[Long-Context Bench](agentic/code/addons/long-context-bench/)** | Benchmark compressed skim plus exact recovery against current context baselines |
858
+ | **[Natural-Language Harness](agentic/code/addons/natural-language-harness/)** | Map inspectable natural-language policy documents to deterministic AIWG mechanisms and ablation reports |
859
+ | **[Premortem v2](agentic/code/addons/premortem-v2/)** | Generate, select, and independently verify bounded risk sets before execution |
860
+ | **[Century Readiness](agentic/code/addons/century-readiness/)** | Review long-horizon stewardship, degradation, replacement, evidence, and meaning-preservation risks |
861
+ | **[Dataset Intelligence](agentic/code/addons/dataset-intelligence/)** | Route dataset intake, planning, materialization, traceability, verification, export, synchronization, and retirement through governed workflows |
862
+ | **[Schema Governance](agentic/code/addons/schema-governance/)** | Discover, author, validate, evolve, and normalize schemas across datasets and SDLC artifacts |
863
+ | **[Compound Memory](agentic/code/addons/compound-memory/)** | Govern promotion from raw evidence and session candidates into line memory or linked wiki knowledge with lineage |
864
+ | **[Line Memory](agentic/code/addons/line-memory/)** | Keep a bounded plain-text set of durable project facts with recency retention and reviewed lifecycle operations |
865
+ | **[LLM Wiki](agentic/code/addons/llm-wiki/)** | Maintain a Markdown wiki topology for entities, concepts, sources, comparisons, and syntheses |
866
+ | **[Semantic Memory](agentic/code/addons/semantic-memory/)** | Provide topology-agnostic memory operations for ingest, lint, query/capture, and event logging |
867
+ | **[Auto Memory](agentic/code/addons/auto-memory/)** | Seed Claude Code Automatic Memory files with AIWG-aware testing, debugging, and architecture sections |
868
+ | **[Agent Persistence](agentic/code/addons/agent-persistence/)** | Supply reusable human-in-the-loop gate definitions for destructive actions, overrides, and recovery escalation |
869
+ | **[AIWG Hooks](agentic/code/addons/aiwg-hooks/)** | Provide hook templates for workflow tracing, permissions, session management, context injection, and quality gates |
870
+ | **[AIWG Fleet](agentic/code/addons/aiwg-fleet/)** | Apply quiet-bot, mention-only participation, and small-plan cost-discipline policies across multi-project fleets |
871
+ | **[Browser Control](agentic/code/addons/browser-control/)** | Drive a user-authorized Chromium-derived browser through Playwright MCP with allow-list and audit boundaries |
872
+ | **[Droid Bridge](agentic/code/addons/droid-bridge/)** | Bridge Claude Code to Factory Droid for batch operations and automated fixes through MCP |
873
+ | **[MCP/UAT Toolkit](agentic/code/addons/uat-mcp/)** | Generate, execute, and report user-acceptance tests against MCP tool surfaces |
874
+ | **[Civic Action](agentic/code/addons/civic-action/)** | Prepare evidence-bound civic research, public-records planning, meeting review, local-resource profiles, corrections, and publication review |
875
+ | **[Network Analysis](agentic/code/addons/network-analysis/)** | Governed saved-PCAP/PCAPNG analysis with bounded TShark recipes, cited packet evidence, and optional local Termshark review |
876
+ | **[Testing Quality](agentic/code/addons/testing-quality/)** | Add TDD gates, mutation testing, flaky-test detection/repair, test-data factories, and test-suite synchronization |
877
+ | **[Writing Quality](agentic/code/addons/writing-quality/)** | Review editorial quality, author requirements, and voice consistency without treating heuristic scores as authorship proof |
878
+ | **[Voice Framework](agentic/code/addons/voice-framework/)** | Define, analyze, blend, and apply reusable writing voice profiles and runtime-selectable output modes |
879
+ | **[Color Palette](agentic/code/addons/color-palette/)** | Generate and review accessible color palettes using color theory, trend research, and WCAG checks |
880
+ | **[Doc Intelligence](agentic/code/addons/doc-intelligence/)** | Scrape, extract, split, audit, and synchronize documentation sources |
881
+ | **[Prose Integration](agentic/code/addons/prose-integration/)** | Detect, read, validate, wire, and run OpenProse contract programs in supported AIWG sessions |
882
+ | **[NLP Prod](agentic/code/addons/nlp-prod/)** | Design and productionize LLM inference pipelines with eval-first, pattern-guided workflow support |
883
+ | **[Context Curator](agentic/code/addons/context-curator/)** | Filter distractors and curate context packs for agent work where irrelevant material can derail results |
884
+ | **[Twelve-Factor](agentic/code/addons/twelve-factor/)** | Review or design applications against Twelve-Factor and modern cloud-native criteria |
885
+ | **[Verbalized Sampling](agentic/code/addons/verbalized-sampling/)** | Apply and evaluate verbalized probability-distribution prompting for output diversity experiments |
886
+ | **[Star Prompt](agentic/code/addons/star-prompt/)** | Offer a tasteful repository-star prompt after successful command completion |
887
+
888
+ Addon details live in each source directory and, where public docs exist, under
889
+ `docs/addons/`. Use [Key Addons](docs/getting-started/key-addons.md) for a
890
+ guided end-user selection path.
729
891
 
730
892
  ---
731
893
 
732
- ### Agents (188)
894
+ ### Agents
733
895
 
734
- Specialized AI personas deployed to your platform with defined tools, responsibilities, and operating rhythms.
896
+ Specialized AI personas deploy to your platform with defined responsibilities,
897
+ tools, and operating rhythms. The exact inventory changes as frameworks evolve,
898
+ so this README keeps durable groupings and examples instead of relying on one
899
+ global total.
735
900
 
736
- #### SDLC Agents (90)
901
+ #### SDLC Agents
737
902
 
738
- | Domain | Agents | Examples |
739
- |--------|--------|---------|
740
- | **Testing & Quality** | 11 | Test Engineer, Test Architect, Mutation Analyst, Regression Analyst, Laziness Detector, Reliability Engineer |
741
- | **Security & Compliance** | 9 | Security Auditor, Security Architect, Compliance Checker, Privacy Officer, Citation Verifier |
742
- | **Architecture & Design** | 12 | Architecture Designer, API Designer, Cloud Architect, System Analyst, Product Designer, Decision Matrix Expert |
743
- | **DevOps & Cloud** | 8 | AWS Specialist, Azure Specialist, GCP Specialist, Kubernetes Expert, DevOps Engineer, Multi-Cloud Strategist |
744
- | **Backend & Data** | 10 | Django Expert, Spring Boot Expert, Data Engineer, Database Optimizer, Software Implementer, Incident Responder |
745
- | **Frontend & Mobile** | 6 | React Expert, Frontend Specialist, Mobile Developer, Accessibility Specialist, UX Lead |
746
- | **AI/ML & Performance** | 5 | AI/ML Engineer, Performance Engineer, Cost Optimizer, Metrics Analyst |
747
- | **Code Quality** | 11 | Code Reviewer, Debugger, Dead Code Analyzer, Technical Debt Analyst, Legacy Modernizer |
748
- | **Documentation** | 7 | Technical Writer, Documentation Synthesizer, Documentation Archivist, Context Librarian |
749
- | **Requirements & Planning** | 7 | Requirements Analyst, Requirements Reviewer, Intake Coordinator, RACI Expert |
750
- | **Agent/Tool Smiths** | 9 | AgentSmith, CommandSmith, MCPSmith, SkillSmith, ToolSmith |
751
- | **Governance & Meta** | 3 | Executive Orchestrator, Recovery Orchestrator, Migration Planner |
903
+ | Domain | Examples |
904
+ |--------|----------|
905
+ | **Testing & Quality** | Test Engineer, Test Architect, Mutation Analyst, Regression Analyst, Reliability Engineer |
906
+ | **Security & Compliance** | Security Auditor, Security Architect, Compliance Checker, Privacy Officer, Citation Verifier |
907
+ | **Architecture & Design** | Architecture Designer, API Designer, Cloud Architect, System Analyst, Product Designer, Decision Matrix Expert |
908
+ | **DevOps & Cloud** | AWS Specialist, Azure Specialist, GCP Specialist, Kubernetes Expert, DevOps Engineer, Multi-Cloud Strategist |
909
+ | **Backend & Data** | Django Expert, Spring Boot Expert, Data Engineer, Database Optimizer, Software Implementer, Incident Responder |
910
+ | **Frontend & Mobile** | React Expert, Frontend Specialist, Mobile Developer, Accessibility Specialist, UX Lead |
911
+ | **AI/ML & Performance** | AI/ML Engineer, Performance Engineer, Cost Optimizer, Metrics Analyst |
912
+ | **Code Quality** | Code Reviewer, Debugger, Dead Code Analyzer, Technical Debt Analyst, Legacy Modernizer |
913
+ | **Documentation** | Technical Writer, Documentation Synthesizer, Documentation Archivist, Context Librarian |
914
+ | **Requirements & Planning** | Requirements Analyst, Requirements Reviewer, Intake Coordinator, RACI Expert |
915
+ | **Agent/Tool Smiths** | AgentSmith, CommandSmith, MCPSmith, SkillSmith, ToolSmith |
916
+ | **Governance & Meta** | Executive Orchestrator, Recovery Orchestrator, Migration Planner |
752
917
 
753
- #### Forensics Agents (13)
918
+ #### Forensics Agents
754
919
 
755
920
  | Agent | What It Does |
756
921
  |-------|-------------|
757
- | Forensics Orchestrator | Coordinates full investigation lifecycle from scoping through reporting |
758
- | Triage Agent | Quick volatile data capture following RFC 3227 volatility order |
759
- | Acquisition Agent | Evidence collection with chain of custody and SHA-256 hash verification |
760
- | Log Analyst | Auth.log, syslog, journal, and application log analysis for brute force, privilege escalation, lateral movement |
761
- | Persistence Hunter | Sweeps cron, systemd, SSH keys, LD_PRELOAD, PAM modules, kernel modules — maps to MITRE ATT&CK |
762
- | Container Analyst | Docker, containerd, Kubernetes forensics — privilege escalation, container escapes, eBPF monitoring |
763
- | Network Analyst | Connection state, DNS, traffic patterns — beaconing, C2, data exfiltration detection |
764
- | Memory Analyst | Volatility 3 memory forensics — process analysis, rootkit detection, credential extraction |
765
- | Cloud Analyst | AWS/Azure/GCP audit logs, IAM review, network flows, API activity anomaly detection |
766
- | Timeline Builder | Multi-source event correlation — chronological incident timelines with attribution |
767
- | IOC Analyst | IOC extraction, enrichment, STIX 2.1 formatting — actionable IOC register |
768
- | Recon Agent | Target reconnaissance — system topology, services, users, network baselines |
769
- | Reporting Agent | Structured forensic reports — executive summary, technical findings, timeline, remediation |
770
-
771
- #### Marketing Agents (37)
772
-
773
- | Domain | Agents |
774
- |--------|--------|
922
+ | Forensics Orchestrator | Coordinates investigation scope, evidence handling, analysis, and reporting |
923
+ | Triage Agent | Captures volatile data following evidence-priority guidance |
924
+ | Acquisition Agent | Collects evidence with chain-of-custody and hash verification |
925
+ | Log Analyst | Reviews auth, syslog, journal, and application logs for suspicious activity |
926
+ | Persistence Hunter | Checks cron, systemd, SSH keys, LD_PRELOAD, PAM modules, and kernel-module indicators |
927
+ | Container Analyst | Reviews Docker, containerd, and Kubernetes evidence |
928
+ | Network Analyst | Reviews connection state, DNS, beaconing, and exfiltration indicators |
929
+ | Memory Analyst | Supports Volatility-style memory forensics workflows |
930
+ | Cloud Analyst | Reviews AWS, Azure, and GCP audit trails and IAM posture |
931
+ | Timeline Builder | Correlates events into chronological incident timelines |
932
+ | IOC Analyst | Extracts and formats indicators for downstream response |
933
+ | Recon Agent | Builds a target baseline for authorized investigation |
934
+ | Reporting Agent | Produces structured executive and technical investigation reports |
935
+
936
+ #### Marketing Agents
937
+
938
+ | Domain | Examples |
939
+ |--------|----------|
775
940
  | **Strategy** | Campaign Strategist, Brand Guardian, Positioning Specialist, Market Researcher, Content Strategist, Channel Strategist |
776
941
  | **Creation** | Copywriter, Content Writer, Email Marketer, Social Media Specialist, SEO Specialist, Graphic Designer, Art Director |
777
942
  | **Management** | Campaign Orchestrator, Production Coordinator, Traffic Manager, Asset Manager, Workflow Coordinator |
778
943
  | **Analytics** | Marketing Analyst, Data Analyst, Attribution Specialist, Reporting Specialist, Budget Planner |
779
944
  | **Communications** | PR Specialist, Crisis Communications, Corporate Communications, Internal Communications, Media Relations |
780
945
 
781
- #### Research Agents (8)
946
+ #### Other Framework Agents
782
947
 
783
- Discovery Agent, Acquisition Agent, Documentation Agent, Citation Agent, Quality Agent, Archival Agent, Provenance Agent, Workflow Agent
784
-
785
- #### Media Curator Agents (6)
786
-
787
- Discography Analyst, Source Discoverer, Acquisition Manager, Quality Assessor, Metadata Curator, Completeness Tracker
948
+ Research uses discovery, acquisition, documentation, citation, quality,
949
+ archival, provenance, and workflow roles. Media Curator uses discography/source,
950
+ acquisition, quality, metadata, and completeness roles. Ops Complete adds
951
+ runbook execution and inventory roles. Security Engineering adds security
952
+ specialists for applied security decisions and supply-chain review.
788
953
 
789
954
  ---
790
955
 
791
- ### Rules (35)
956
+ ### Rules
792
957
 
793
- Enforcement patterns that prevent common AI failure modes. Rules deploy automatically with their framework.
958
+ Rules are durable guardrails that deploy with the frameworks or addons that own
959
+ them. They prevent common agent failure modes and define review boundaries.
794
960
 
795
- #### Core Rules (10) — Always Active
961
+ #### Core Rules
796
962
 
797
963
  | Rule | Severity | What It Enforces |
798
964
  |------|----------|-----------------|
799
- | `no-attribution` | CRITICAL | AI tools are tools — never add attribution to commits, PRs, docs, or code |
800
- | `token-security` | CRITICAL | Never hard-code tokens; use heredoc pattern for scoped lifetime; file permissions 600 |
801
- | `versioning` | CRITICAL | CalVer YYYY.M.PATCH with NO leading zeros; npm rejects leading zeros |
802
- | `citation-policy` | CRITICAL | Never fabricate citations, DOIs, or URLs; only cite verified sources; GRADE-appropriate hedging |
803
- | `anti-laziness` | HIGH | Never delete tests to pass, skip tests, remove features, or weaken assertions; escalate after 3 failures |
804
- | `executable-feedback` | HIGH | Execute tests before returning code; track execution history; max 3 retries with root cause analysis |
805
- | `failure-mitigation` | HIGH | Detect and recover from 6 LLM failure archetypes: hallucination, context loss, instruction drift, safety, technical, consistency |
806
- | `research-before-decision` | HIGH | Research codebase before acting: IDENTIFY → SEARCH → EXTRACT → REASON → ACT → VERIFY |
807
- | `instruction-comprehension` | HIGH | Fully parse all instructions before acting; track multi-part requests to completion |
808
- | `subagent-scoping` | HIGH | One focused task per subagent; <20% context budget; no delegation chains deeper than 2 levels |
809
-
810
- #### SDLC Rules (34) — Active with Framework
811
-
812
- Actionable feedback, mention wiring, HITL gates, agent fallback, provenance tracking, TAO loop, reproducibility validation, SDLC orchestration, agent-friendly code, agent generation guardrails, artifact discovery, HITL patterns, human gate display, thought protocol, reasoning sections, few-shot examples, best output selection, reproducibility, progressive disclosure, conversable agent interface, auto-reply chains, criticality panel sizing, qualified references.
813
-
814
- #### Research Rules (2) — Active with Research
815
-
816
- Research metadata (FAIR-compliant YAML frontmatter), index generation (auto-generated INDEX.md per FAIR F4).
965
+ | `no-attribution` | CRITICAL | AI tools are tools; do not add AI attribution to commits, PRs, docs, or code |
966
+ | `token-security` | CRITICAL | Keep tokens and secrets out of source; use scoped lifetime and restricted file permissions |
967
+ | `versioning` | CRITICAL | Use the repository's CalVer release format consistently |
968
+ | `citation-policy` | CRITICAL | Do not fabricate citations, DOIs, URLs, or research claims |
969
+ | `anti-laziness` | HIGH | Do not delete tests, skip required checks, remove features, or weaken assertions to pass |
970
+ | `executable-feedback` | HIGH | Run appropriate validation before returning implementation work |
971
+ | `failure-mitigation` | HIGH | Detect and recover from hallucination, context loss, instruction drift, safety, technical, and consistency failures |
972
+ | `research-before-decision` | HIGH | Inspect the codebase and docs before making technical decisions |
973
+ | `instruction-comprehension` | HIGH | Parse prohibitions, requirements, and preferences before acting |
974
+ | `subagent-scoping` | HIGH | Keep delegated tasks focused and bounded when delegation is used |
975
+
976
+ #### Domain Rules
977
+
978
+ | Domain | Examples |
979
+ |--------|----------|
980
+ | **SDLC** | HITL gates, provenance tracking, artifact discovery, phase gates, reproducibility validation, agent-friendly code, fallback, review, and handoff rules |
981
+ | **Forensics** | Evidence integrity, chain of custody, forensic reporting, and authorized investigation boundaries |
982
+ | **Security Engineering** | Cryptographic decision boundaries, runtime secret hygiene, supply-chain trust, physical-access threat modeling, and DFIR readiness handoff |
983
+ | **Ops** | Ops safety, executable runbook format, evidence governance, issue tracking, and cross-repo reference rules |
984
+ | **Civic Action** | Human authority, citation, publication, public-source, privacy, and anti-targeting boundaries |
985
+ | **Addon Rules** | Browser authorization, dataset boundaries, agentic installer safety, voice/output behavior, and hook discipline |
817
986
 
818
987
  ---
819
988
 
820
- ### Skills (128)
821
-
822
- Natural language workflow triggers. Say "what's the project status?" and the `project-awareness` skill activates.
823
-
824
- | Category | Skills | Examples |
825
- |----------|--------|---------|
826
- | **Regression Testing** | 12 | regression-check, regression-baseline, regression-bisect, regression-performance, regression-api-contract, regression-cicd-hooks, regression-learning |
827
- | **Voice & Writing** | 6 | voice-create, voice-analyze, voice-apply, voice-blend, ai-pattern-detection, brand-compliance |
828
- | **Testing & Quality** | 8 | auto-test-execution, test-coverage, test-sync, mutation-test, flaky-detect, flaky-fix, tdd-enforce, qa-protocol |
829
- | **Forensics & Security** | 8 | linux-forensics, memory-forensics, cloud-forensics, container-forensics, sigma-hunting, log-analysis, ioc-extraction, supply-chain-forensics |
830
- | **SDLC & Workflow** | 10 | sdlc-accelerate, sdlc-reports, gate-evaluation, approval-workflow, iteration-control, risk-cycle, parallel-dispatch, decision-support |
831
- | **Documentation** | 6 | doc-sync, doc-scraper, doc-splitter, llms-txt-support, pdf-extractor, source-unifier |
832
- | **Artifacts & Traceability** | 6 | artifact-orchestration, artifact-metadata, artifact-lookup, traceability-check, claims-validator, citation-guard |
833
- | **Research** | 2 | grade-on-ingest, auto-provenance |
834
- | **Infrastructure** | 5 | config-validator, template-engine, code-chunker, decompose-file, workspace-health |
835
- | **Iteration** | 4 | agent-loop, issue-driven-ralph, cross-task-learner, reflection-injection |
836
- | **Other** | 19 | performance-digest, competitive-intel, audience-synthesis, skill-builder, skill-enhancer, skill-packager, quality-checker, nl-router, tot-exploration, and more |
989
+ ### Skills
990
+
991
+ Skills are natural-language workflows. A user describes an outcome, the agent
992
+ discovers the relevant skill, loads its instructions, and applies its protocol.
993
+ The current repo contains a large and changing skill surface, so this section
994
+ keeps durable categories and examples.
995
+
996
+ | Category | Examples |
997
+ |----------|---------|
998
+ | **Capability discovery and setup** | `aiwg-utils-quickref`, `steward`, `aiwg-status`, `aiwg-doctor`, `use`, provider regeneration |
999
+ | **SDLC and delivery** | `intake-wizard`, `sdlc-accelerate`, gate evaluation, delivery-track flows, deployment, guided implementation |
1000
+ | **Testing and quality** | `tdd-enforce`, `mutation-test`, `flaky-detect`, `flaky-fix`, `test-sync`, factory generation |
1001
+ | **Security and forensics** | supply-chain hardening, auth-factor design, degraded-mode review, DFIR readiness, log analysis, IOC extraction |
1002
+ | **Research and knowledge** | source acquisition, paper induction, GRADE checks, citation work, wiki ingest, synthesis, knowledge-base health |
1003
+ | **Marketing and content** | campaign intake, creative brief, brand compliance, social strategy, email campaigns, performance digests |
1004
+ | **Media curation** | source discovery, acquisition planning, transcript sidecars, metadata tagging, quality filtering, archive verification |
1005
+ | **Datasets and schemas** | dataset intake, source assessment, capability recommendation, plan review, ingest, trace, verify, export, retire |
1006
+ | **Memory and persistence** | line-memory operations, compound-memory review, semantic-memory capture/query, llm-wiki topology |
1007
+ | **Operations and automation** | runbook execution, ops verification, activity logs, hooks, daemon sessions, schedule support |
1008
+ | **Authoring and development** | skill creation, addon/framework scaffolding, validation, schema governance, doc synchronization |
837
1009
 
838
1010
  ---
839
1011
 
@@ -841,9 +1013,11 @@ Natural language workflow triggers. Say "what's the project status?" and the `pr
841
1013
 
842
1014
  ### SDLC Complete — Full Software Development Lifecycle
843
1015
 
844
- The SDLC framework implements a phase-gated development lifecycle with 90 specialized agents, 34 enforcement rules, and 170+ artifact templates. Natural language commands drive phase transitions with automated quality gates.
1016
+ The SDLC framework implements a phase-gated development lifecycle with
1017
+ specialized agents, enforcement rules, and artifact templates. Natural-language
1018
+ requests drive phase transitions with reviewable quality gates.
845
1019
 
846
- ```
1020
+ ```text
847
1021
  ┌──────────┐ ┌─────────────┐ ┌──────────────┐ ┌────────────┐ ┌────────────┐
848
1022
  │ CONCEPT │───▶│ INCEPTION │───▶│ ELABORATION │───▶│CONSTRUCTION│───▶│ TRANSITION │
849
1023
  │ │ │ │ │ │ │ │ │ │
@@ -854,7 +1028,7 @@ The SDLC framework implements a phase-gated development lifecycle with 90 specia
854
1028
  └──────────┘ └──────┬──────┘ └──────┬───────┘ └─────┬──────┘ └────────────┘
855
1029
  │ │ │
856
1030
  ┌──▼──┐ ┌──▼──┐ ┌──▼──┐
857
- │ LOM │ │ ABM │ │ IOC │ ← Quality Gates
1031
+ │ LOM │ │ ABM │ │ IOC │
858
1032
  │Gate │ │Gate │ │Gate │
859
1033
  └─────┘ └─────┘ └─────┘
860
1034
 
@@ -862,36 +1036,36 @@ The SDLC framework implements a phase-gated development lifecycle with 90 specia
862
1036
  IOC = Initial Operational Capability
863
1037
  ```
864
1038
 
865
- **SDLC Flow Commands (24):**
1039
+ **SDLC Flow Commands:**
866
1040
 
867
1041
  | Command | Phase | What It Does |
868
1042
  |---------|-------|-------------|
869
- | `/intake-wizard` | Concept | Generate project intake form from natural language description |
870
- | `/intake-start` | Concept→Inception | Validate intake, kick off with agent assignments |
871
- | `/intake-from-codebase` | Concept | Scan existing codebase, generate intake from analysis |
872
- | `/flow-concept-to-inception` | Concept→Inception | Phase transition with intake validation and vision alignment |
873
- | `/flow-inception-to-elaboration` | Inception→Elaboration | Architecture baselining and risk retirement |
874
- | `/flow-elaboration-to-construction` | Elaboration→Construction | Iteration planning, team scaling, full-scale development |
875
- | `/flow-construction-to-transition` | Construction→Transition | IOC validation, production deployment, operational handover |
876
- | `/flow-discovery-track` | Any | Prepare validated requirements one iteration ahead of delivery |
877
- | `/flow-delivery-track` | Any | Test-driven development, quality gates, iteration assessment |
878
- | `/flow-iteration-dual-track` | Any | Synchronized Discovery + Delivery workflows |
879
- | `/flow-deploy-to-production` | Transition | Strategy selection, validation, automated rollback, regression gates |
880
- | `/flow-incident-response` | Operations | Triage, escalation, resolution, post-incident review (ITIL) |
881
- | `/flow-security-review-cycle` | Any | Continuous security validation, threat modeling, vulnerability management |
882
- | `/flow-performance-optimization` | Any | Baseline, bottleneck ID, optimization, load testing, SLO validation |
883
- | `/flow-retrospective-cycle` | Any | Structured feedback, improvement tracking, action items |
884
- | `/flow-change-control` | Any | Baseline management, impact assessment, CCB review, communication |
885
- | `/flow-risk-management-cycle` | Any | Continuous risk identification, assessment, tracking, retirement |
886
- | `/flow-compliance-validation` | Any | Requirements mapping, audit evidence, gap analysis, attestation |
887
- | `/flow-knowledge-transfer` | Transition | Assessment, documentation, shadowing, validation, handover |
888
- | `/flow-team-onboarding` | Any | Pre-boarding, training, buddy assignment, 30/60/90 day check-ins |
889
- | `/flow-hypercare-monitoring` | Transition | 24/7 support, SLO tracking, rapid issue response |
890
- | `/flow-gate-check` | Any | Multi-agent phase gate validation with comprehensive reporting |
891
- | `/flow-handoff-checklist` | Any | Handoff validation between phases and tracks |
892
- | `/flow-guided-implementation` | Construction | Bounded iteration with issue-to-code automation |
893
-
894
- **SDLC Accelerate — Idea to Construction-Ready in One Command:**
1043
+ | `/intake-wizard` | Concept | Generate project intake from a natural-language description |
1044
+ | `/intake-start` | Concept -> Inception | Validate intake and begin agent assignments |
1045
+ | `/intake-from-codebase` | Concept | Scan an existing codebase and generate intake from analysis |
1046
+ | `/flow-concept-to-inception` | Concept -> Inception | Transition with intake validation and vision alignment |
1047
+ | `/flow-inception-to-elaboration` | Inception -> Elaboration | Baseline architecture and retire major risks |
1048
+ | `/flow-elaboration-to-construction` | Elaboration -> Construction | Prepare iteration planning, scale delivery, and begin implementation |
1049
+ | `/flow-construction-to-transition` | Construction -> Transition | Validate IOC, deployment readiness, and operational handoff |
1050
+ | `/flow-discovery-track` | Any | Prepare validated requirements ahead of delivery |
1051
+ | `/flow-delivery-track` | Any | Run test-driven delivery with quality gates |
1052
+ | `/flow-iteration-dual-track` | Any | Coordinate discovery and delivery tracks |
1053
+ | `/flow-deploy-to-production` | Transition | Select deployment strategy, validate, and prepare rollback/regression checks |
1054
+ | `/flow-incident-response` | Operations | Triage, resolve, and review incidents |
1055
+ | `/flow-security-review-cycle` | Any | Run continuous security validation and threat review |
1056
+ | `/flow-performance-optimization` | Any | Baseline, identify bottlenecks, optimize, and validate SLOs |
1057
+ | `/flow-retrospective-cycle` | Any | Capture feedback and track improvement actions |
1058
+ | `/flow-change-control` | Any | Assess impact, coordinate review, and manage communication |
1059
+ | `/flow-risk-management-cycle` | Any | Identify, assess, track, and retire risks |
1060
+ | `/flow-compliance-validation` | Any | Map requirements, collect evidence, and identify gaps |
1061
+ | `/flow-knowledge-transfer` | Transition | Prepare documentation, shadowing, validation, and handover |
1062
+ | `/flow-team-onboarding` | Any | Structure onboarding, training, buddy support, and follow-up |
1063
+ | `/flow-hypercare-monitoring` | Transition | Track early-life support, SLOs, and rapid-response items |
1064
+ | `/flow-gate-check` | Any | Run multi-agent phase-gate validation |
1065
+ | `/flow-handoff-checklist` | Any | Validate handoff between phases and tracks |
1066
+ | `/flow-guided-implementation` | Construction | Run bounded issue-to-code iteration with validation and escalation |
1067
+
1068
+ **SDLC Accelerate — from idea to reviewed planning artifacts:**
895
1069
 
896
1070
  ```bash
897
1071
  # From a description
@@ -904,11 +1078,13 @@ aiwg sdlc-accelerate --from-codebase .
904
1078
  aiwg sdlc-accelerate --resume
905
1079
  ```
906
1080
 
907
- Generates intake form, vision document, use cases, architecture baseline, risk register, test strategy, and deployment plan — all with human approval gates between phases.
1081
+ It can generate intake, vision, use cases, architecture baseline, risk register,
1082
+ test strategy, and deployment planning artifacts with human review between
1083
+ major phases.
908
1084
 
909
1085
  **Dual-Track Iteration Model:**
910
1086
 
911
- ```
1087
+ ```text
912
1088
  ┌─────────────────────────────────────────────────┐
913
1089
  │ ITERATION N │
914
1090
  │ │
@@ -931,21 +1107,23 @@ Generates intake form, vision document, use cases, architecture baseline, risk r
931
1107
  └─────────────────────────────────────────────────┘
932
1108
  ```
933
1109
 
934
- **Metrics & Quality Tracking:**
1110
+ **Metrics and Quality Tracking:**
935
1111
 
936
1112
  | Metric Category | Metrics Tracked |
937
1113
  |-----------------|-----------------|
938
- | **DORA** (4) | Deployment Frequency, Lead Time, Change Failure Rate, MTTR |
939
- | **Velocity** (3) | Story Points, Cycle Time, Throughput |
940
- | **Flow** (3) | WIP Limits, Flow Efficiency, Blocked Items |
941
- | **Quality** (13) | Test Coverage (4), Defect Metrics (4), Code Quality (3), Technical Debt (2) |
942
- | **Operational** (16) | SLO/SLI (5), Infrastructure (4), Incidents (4), Cost (3) |
1114
+ | **DORA** | Deployment frequency, lead time, change failure rate, MTTR |
1115
+ | **Velocity** | Story points, cycle time, throughput |
1116
+ | **Flow** | WIP limits, flow efficiency, blocked items |
1117
+ | **Quality** | Test coverage, defect metrics, code quality, technical debt |
1118
+ | **Operational** | SLO/SLI, infrastructure, incidents, cost |
943
1119
 
944
- ### Forensics Complete — Digital Forensics & Incident Response
1120
+ ### Forensics Complete — Digital Forensics and Incident Response
945
1121
 
946
- Full DFIR investigation workflow following NIST SP 800-86, with MITRE ATT&CK mapping and Sigma rule hunting.
1122
+ Forensics Complete supports authorized DFIR work following NIST SP 800-86-style
1123
+ evidence handling, MITRE ATT&CK mapping, Sigma hunting, timeline construction,
1124
+ and structured reporting.
947
1125
 
948
- ```
1126
+ ```text
949
1127
  ┌──────────┐ ┌──────────┐ ┌────────────┐ ┌──────────┐ ┌──────────┐
950
1128
  │ SCOPE │───▶│ TRIAGE │───▶│ ACQUIRE │───▶│ ANALYZE │───▶│ REPORT │
951
1129
  │ │ │ │ │ │ │ │ │ │
@@ -962,29 +1140,16 @@ Full DFIR investigation workflow following NIST SP 800-86, with MITRE ATT&CK map
962
1140
  **Investigation Commands:**
963
1141
 
964
1142
  ```bash
965
- /forensics-profile # Build target system profile via SSH
966
- /forensics-triage # Quick triage following RFC 3227 volatility order
967
- /forensics-acquire # Evidence acquisition with chain of custody
968
- /forensics-investigate # Full multi-agent investigation workflow
969
- /forensics-timeline # Build correlated event timeline
970
- /forensics-hunt # Threat hunt using Sigma rules
971
- /forensics-ioc # Extract and enrich IOCs
972
- /forensics-report # Generate forensic investigation report
973
- /forensics-status # Show investigation dashboard
974
- ```
975
-
976
- **Bundled Sigma Rules (8):**
977
-
978
- | Rule | What It Detects |
979
- |------|----------------|
980
- | SSH Brute Force | Repeated failed SSH authentication attempts |
981
- | Unauthorized SUID | Unexpected SUID/SGID binaries |
982
- | LD_PRELOAD Rootkit | Library injection via LD_PRELOAD |
983
- | Cron Persistence | Unauthorized crontab modifications |
984
- | Kernel Module Load | Suspicious kernel module insertion |
985
- | PAM Backdoor | PAM configuration tampering |
986
- | SSH Key Injection | Unauthorized authorized_keys modifications |
987
- | Systemd Persistence | Suspicious systemd unit creation |
1143
+ /forensics-profile
1144
+ /forensics-triage
1145
+ /forensics-acquire
1146
+ /forensics-investigate
1147
+ /forensics-timeline
1148
+ /forensics-hunt
1149
+ /forensics-ioc
1150
+ /forensics-report
1151
+ /forensics-status
1152
+ ```
988
1153
 
989
1154
  **Supported Evidence Sources:**
990
1155
 
@@ -992,49 +1157,70 @@ Full DFIR investigation workflow following NIST SP 800-86, with MITRE ATT&CK map
992
1157
  |--------|-------|----------|
993
1158
  | Auth logs | Log Analyst | Brute force, privilege escalation, lateral movement |
994
1159
  | Syslog / journal | Log Analyst | System events, service anomalies |
995
- | Network connections | Network Analyst | C2 beaconing, data exfiltration, DNS tunneling |
996
- | Docker/containerd | Container Analyst | Container escapes, image tampering, eBPF monitoring |
997
- | Memory dumps | Memory Analyst | Process injection, rootkits, credential extraction |
998
- | AWS/Azure/GCP | Cloud Analyst | API anomalies, IAM abuse, network flow analysis |
1160
+ | Network connections | Network Analyst | C2 beaconing, exfiltration, DNS tunneling |
1161
+ | Docker/containerd | Container Analyst | Container escape, image tampering, runtime evidence |
1162
+ | Memory dumps | Memory Analyst | Process analysis, rootkits, credential artifacts |
1163
+ | AWS/Azure/GCP | Cloud Analyst | API anomalies, IAM abuse, network-flow evidence |
999
1164
  | File system | Persistence Hunter | Cron, systemd, SSH keys, PAM, kernel modules |
1000
1165
 
1001
1166
  ### Media/Marketing Kit — Campaign Lifecycle
1002
1167
 
1003
- ```
1168
+ Media/Marketing Kit treats campaign work as a lifecycle with artifacts and
1169
+ review gates, so strategy, content, legal/brand review, publication planning,
1170
+ and performance analysis remain inspectable.
1171
+
1172
+ ```text
1004
1173
  ┌──────────┐ ┌──────────┐ ┌──────────┐ ┌──────────┐ ┌──────────┐
1005
1174
  │ STRATEGY │───▶│ CREATION │───▶│ REVIEW │───▶│ PUBLISH │───▶│ ANALYZE │
1006
1175
  │ │ │ │ │ │ │ │ │ │
1007
1176
  │ Research │ │ Copy │ │ Brand │ │ Schedule │ │ KPIs │
1008
1177
  │ Audience │ │ Design │ │ Legal │ │ Channels │ │ Reports │
1009
- │ Strategy │ │ Content │ │ Quality │ │ Launch │ │ ROI │
1178
+ │ Strategy │ │ Content │ │ Quality │ │ Launch │ │ Learnings│
1010
1179
  └──────────┘ └──────────┘ └──────────┘ └──────────┘ └──────────┘
1011
1180
  ```
1012
1181
 
1013
- 37 agents across strategy, creation, management, analytics, and communications. 87+ templates covering campaign intake, brand guidelines, content briefs, social playbooks, email sequences, PR kits, and analytics dashboards.
1182
+ | Discipline | Example Artifacts |
1183
+ |------------|-------------------|
1184
+ | Strategy | Campaign intake, positioning, messaging, audience profile, channel plan |
1185
+ | Creation | Blog drafts, social posts, email sequences, creative briefs, media kits |
1186
+ | Review | Brand compliance, legal clearance, accessibility review, claim substantiation |
1187
+ | Publication | Launch checklist, schedule, channel handoff, go-live readiness |
1188
+ | Analysis | KPI report, performance digest, retrospective, optimization plan |
1189
+
1190
+ ### Media Curator — Archive Management
1014
1191
 
1015
- ### Media Curator — Intelligent Archive Management
1192
+ Media Curator helps assess, acquire, organize, verify, transcribe, and export
1193
+ media collections. It starts with assessment and planning so unknown or mixed
1194
+ media is routed before downloads or metadata rewrites.
1016
1195
 
1017
1196
  ```bash
1018
1197
  # Full curation pipeline
1019
1198
  /curate "Pink Floyd"
1020
1199
 
1021
1200
  # Step by step
1022
- /analyze-artist "Pink Floyd" # Identify eras, catalog structure
1023
- /find-sources "Pink Floyd" "DSOTM" # Discover across YouTube, Archive.org, Bandcamp
1024
- /acquire # Download with format selection
1025
- /transcribe-media /path/to/media.wav # Create timestamped transcript sidecars
1026
- /tag-collection # Apply metadata, embed artwork, rename
1027
- /check-completeness # Gap analysis against canonical discography
1028
- /assemble "Pink Floyd live 1973" # Build thematic compilations
1029
- /export --format plex # Export to Plex, Jellyfin, MPD, or archival
1030
- /verify-archive # SHA-256 integrity verification
1031
- ```
1032
-
1033
- Quality tiers: Tier 1 (Official/Lossless) → Tier 2 (High Quality) → Tier 3 (Acceptable) → Tier 4 (Avoid). Transcript sidecars preserve source hashes, transcript hashes, timestamps, and optional speaker labels for review and future research handoff. Standards: ID3v2.4, Vorbis Comments, MusicBrainz, PREMIS 3.0, W3C PROV-O.
1201
+ /analyze-artist "Pink Floyd"
1202
+ /find-sources "Pink Floyd" "DSOTM"
1203
+ /acquire
1204
+ /transcribe-media /path/to/media.wav
1205
+ /tag-collection
1206
+ /check-completeness
1207
+ /assemble "Pink Floyd live 1973"
1208
+ /export --format plex
1209
+ /verify-archive
1210
+ ```
1211
+
1212
+ Quality tiers help reviewers choose what to keep. Transcript sidecars preserve
1213
+ source hashes, transcript hashes, timestamps, and optional speaker labels for
1214
+ review and later research handoff. Common standards include ID3v2.4, Vorbis
1215
+ Comments, MusicBrainz, PREMIS 3.0, and W3C PROV-O.
1034
1216
 
1035
1217
  ### Research Complete — Academic Research Pipeline
1036
1218
 
1037
- ```
1219
+ Research Complete turns search results and PDFs into source-grounded,
1220
+ reviewable research artifacts with persistent identifiers, quality checks, and
1221
+ provenance.
1222
+
1223
+ ```text
1038
1224
  ┌──────────┐ ┌──────────┐ ┌──────────┐ ┌──────────┐
1039
1225
  │ DISCOVER │───▶│ ACQUIRE │───▶│ DOCUMENT │───▶│ ARCHIVE │
1040
1226
  │ │ │ │ │ │ │ │
@@ -1045,20 +1231,92 @@ Quality tiers: Tier 1 (Official/Lossless) → Tier 2 (High Quality) → Tier 3 (
1045
1231
  └──────────┘ └──────────┘ └──────────┘ └──────────┘
1046
1232
  ```
1047
1233
 
1048
- 8-stage pipeline: Discovery → Acquisition → Documentation → Citation → Quality Assessment → Synthesis → Gap Analysis → Archival. Persistent REF-XXX identifiers. GRADE scoring (HIGH/MODERATE/LOW/VERY LOW). Unpaywall integration for open access papers.
1234
+ Pipeline stages: Discovery -> Acquisition -> Documentation -> Citation ->
1235
+ Quality Assessment -> Synthesis -> Gap Analysis -> Archival. The framework uses
1236
+ `REF-XXX` identifiers, GRADE-style evidence quality labels, FAIR-style checks,
1237
+ and Unpaywall lookup for open-access discovery. It flags unsupported claims for
1238
+ review instead of promising error-free summaries.
1239
+
1240
+ ### Knowledge Base — Linked Project Wiki
1241
+
1242
+ Knowledge Base is for open-ended knowledge accumulation where the taxonomy
1243
+ emerges over time. It uses entity, concept, source, comparison, and synthesis
1244
+ pages so future sessions can find what is known, what is missing, and how ideas
1245
+ connect.
1246
+
1247
+ | Page Type | Purpose |
1248
+ |-----------|---------|
1249
+ | Entity | A person, company, tool, place, system, or other named thing |
1250
+ | Concept | A technique, pattern, framework, or idea |
1251
+ | Source | The evidence layer for claims and summaries |
1252
+ | Comparison | A decision aid for tools, approaches, vendors, or options |
1253
+ | Synthesis | A higher-level claim produced by combining multiple sources |
1254
+
1255
+ ### Ops Complete — Executable Operations
1256
+
1257
+ Ops Complete gives operational procedures a structured envelope: inventory,
1258
+ capabilities, playbooks, gates, targets, schedules, pipelines, and extensions.
1259
+ It is useful when procedures must be idempotent, verifiable, and evidence-aware.
1260
+
1261
+ ```yaml
1262
+ apiVersion: ops.aiwg.io/v1
1263
+ kind: OpsPlaybook
1264
+ metadata:
1265
+ name: deploy-auth-stack
1266
+ namespace: production
1267
+ spec:
1268
+ # Desired state
1269
+ status:
1270
+ # Observed state written by the executor
1271
+ ```
1272
+
1273
+ Extensions add domain-specific ops support for systems, IT, development
1274
+ infrastructure, and streaming workflows. See [Ops Complete overview](docs/frameworks/ops-complete/overview.md)
1275
+ and [Ops evidence governance](https://github.com/jmagly/aiwg/blob/main/docs/ops-evidence-governance.md).
1276
+
1277
+ ### Security Engineering — Applied Security Decisions
1278
+
1279
+ Security Engineering complements SDLC and forensics by focusing on security
1280
+ decisions that need explicit assumptions and reviewable tradeoffs.
1281
+
1282
+ | Area | Example Use |
1283
+ |------|-------------|
1284
+ | Cryptographic primitives | Choose AEAD, KDF, hashing, randomness, and signing patterns for a concrete workload |
1285
+ | Chain of trust | Map trust anchors, update paths, verification points, and failure modes |
1286
+ | Authentication factors | Decide factor mix, enrollment, recovery, lockout, and degraded-mode behavior |
1287
+ | Runtime secret hygiene | Review secret storage, process boundaries, rotation, logging, and local development exposure |
1288
+ | Supply-chain trust | Review dependency sources, lifecycle scripts, release provenance, SBOMs, and signed artifacts |
1289
+ | Physical-access threats | Model device seizure, kiosk, lab, field, and hostile-local-user conditions |
1290
+ | DFIR readiness | Prepare evidence handoff points before an incident occurs |
1291
+
1292
+ Install with:
1293
+
1294
+ ```bash
1295
+ aiwg use security-engineering
1296
+ ```
1297
+
1298
+ ### Validation Complete — Focused Validation
1299
+
1300
+ Validation Complete provides a small validation workflow surface for teams that
1301
+ need structured review without adopting a broader lifecycle. Use it where a
1302
+ project already has its own process but wants AIWG-style validation gates and
1303
+ reports.
1049
1304
 
1050
1305
  ---
1051
1306
 
1052
1307
  ## Voice Framework — Content Voice Consistency
1053
1308
 
1054
- 4 built-in voice profiles with create, analyze, blend, and apply skills:
1309
+ Voice Framework defines reusable writing profiles and output modes that can be
1310
+ applied across docs, release notes, campaigns, reports, and internal guidance.
1311
+ It describes the desired voice directly rather than relying only on banned word
1312
+ lists.
1055
1313
 
1056
1314
  | Profile | When to Use | Characteristics |
1057
1315
  |---------|-------------|----------------|
1058
- | `technical-authority` | API docs, architecture guides | Precise terminology, confident assertions, specific metrics |
1059
- | `friendly-explainer` | Tutorials, onboarding | Accessible language, analogies, encouragement |
1060
- | `executive-brief` | Status reports, proposals | Bottom-line-first, quantified impact, action-oriented |
1061
- | `casual-conversational` | Blog posts, social media | Natural rhythm, opinions, varied structure |
1316
+ | `technical-authority` | API docs, architecture guides | Precise terminology, direct claims, concrete examples |
1317
+ | `friendly-explainer` | Tutorials, onboarding | Accessible language, patient sequencing, light warmth |
1318
+ | `executive-brief` | Status reports, proposals | Decision-oriented summaries, concise evidence, clear next steps |
1319
+ | `casual-conversational` | Blog posts, social media | Natural rhythm, opinion-forward phrasing, varied structure |
1062
1320
 
1063
1321
  ```bash
1064
1322
  # Apply a voice to content
@@ -1072,626 +1330,806 @@ Quality tiers: Tier 1 (Official/Lossless) → Tier 2 (High Quality) → Tier 3 (
1072
1330
 
1073
1331
  # Blend two voices
1074
1332
  /voice-blend technical-authority casual-conversational --ratio 70:30
1075
-
1076
- # Detect AI patterns and suggest authentic alternatives
1077
- /ai-pattern-detection docs/generated-content.md
1078
1333
  ```
1079
1334
 
1080
- ---
1335
+ See [Voice Framework overview](docs/addons/voice-framework/overview.md) and
1336
+ [Voice Framework quickstart](docs/addons/voice-framework/quickstart.md).
1081
1337
 
1082
1338
  ## MCP Server — Model Context Protocol Integration
1083
1339
 
1084
- AIWG includes a built-in MCP server for tool-based AI workflow integration:
1340
+ AIWG can expose its project context, discovery catalog, and governed workflows through a Model Context Protocol
1341
+ server. This lets MCP-capable tools call AIWG without learning the repository layout or memorizing provider-specific
1342
+ file locations.
1343
+
1344
+ The MCP server is useful when you want an assistant to ask AIWG questions such as “what capabilities are available for
1345
+ release planning?”, “show the SDLC quickstart”, or “run this governed workflow and return the evidence artifact.” The
1346
+ base server keeps a small default tool surface; larger toolsets can be enabled explicitly for teams that want richer
1347
+ orchestration, mission, dataset, or framework operations.
1085
1348
 
1086
1349
  ```bash
1087
- # Start MCP server
1350
+ # Run the local MCP server
1088
1351
  aiwg mcp serve
1089
1352
 
1090
- # Install into Claude Desktop
1091
- aiwg mcp install claude
1353
+ # Enable additional toolsets for a richer host integration
1354
+ aiwg mcp serve --toolsets=flows,missions,catalog
1355
+
1356
+ # Use an environment variable when the host launches the server command
1357
+ AIWG_MCP_TOOLSETS=flows,missions,catalog aiwg mcp serve
1092
1358
 
1093
- # Show capabilities
1359
+ # Inspect server metadata and supported install targets
1094
1360
  aiwg mcp info
1095
1361
  ```
1096
1362
 
1097
- The MCP server exposes AIWG's artifact management, workflow execution, and project health capabilities as tools that any MCP-compatible AI platform can invoke programmatically.
1363
+ Provider installation depends on the host. AIWG can write MCP configuration for supported targets where the provider
1364
+ has a stable local MCP config format; other providers use the same server command in their own UI or settings.
1098
1365
 
1099
- ---
1366
+ ```bash
1367
+ # Install AIWG MCP config for a supported local target
1368
+ aiwg mcp install claude
1369
+ aiwg mcp install cursor
1370
+ aiwg mcp install codex
1371
+ ```
1372
+
1373
+ MCP integration does not make every AIWG operation model-backed. Catalog reads, status checks, link resolution, and
1374
+ local evidence inspection are ordinary local operations. Workflows that ask an assistant to reason, draft, call
1375
+ another provider, or continue a Ralph loop may use model calls depending on the connected host and selected provider.
1376
+
1377
+ See also: [MCP server documentation](docs/mcp/README.md), [MCP capability
1378
+ audit](docs/integrations/mcp-capability-audit.md), and [cross-platform
1379
+ overview](docs/integrations/cross-platform-overview.md).
1100
1380
 
1101
1381
  ## Agent Evaluation Framework
1102
1382
 
1103
- Test agent quality with archetype resistance testing based on Roig (2025) failure patterns:
1383
+ AIWG treats agents, skills, commands, and rules as reviewable project assets. The evaluation workflow is designed to
1384
+ answer concrete questions before you rely on a capability in a live project:
1385
+
1386
+ - Does the capability declare the right trigger conditions and boundaries?
1387
+ - Does it cite the files, schemas, or rules it depends on?
1388
+ - Does it produce artifacts that another provider can inspect?
1389
+ - Does it fail safely when prerequisites are missing?
1390
+ - Does it preserve evidence for audit, handoff, or regression review?
1391
+
1392
+ For direct CLI checks, use the catalog, metadata validation, skill linting, and evidence commands that are available
1393
+ in the current CLI surface.
1104
1394
 
1105
1395
  ```bash
1106
- # Evaluate a specific agent
1107
- /eval-agent security-auditor
1396
+ # Find relevant evaluation or review capabilities
1397
+ aiwg discover "agent evaluation" --limit 5
1398
+ aiwg discover "skill lint evidence" --limit 5
1399
+
1400
+ # Inspect a selected capability before applying it
1401
+ aiwg show skill aiwg-doctor
1402
+ aiwg show skill context-firewall
1108
1403
 
1109
- # Test archetype resistance
1110
- /eval-agent test-engineer --category archetype
1404
+ # Validate local metadata and skill packaging
1405
+ aiwg validate-metadata
1406
+ aiwg skill-lint
1111
1407
 
1112
- # Run performance benchmarks
1113
- /eval-agent code-reviewer --category performance
1408
+ # Capture and verify evidence bundles when a workflow supports them
1409
+ aiwg evidence export --help
1410
+ aiwg evidence verify --help
1114
1411
  ```
1115
1412
 
1116
- **Test Categories:**
1413
+ A practical evaluation usually starts with a small task and a pass/fail criterion. For example:
1117
1414
 
1118
- | Category | Tests | What It Validates |
1119
- |----------|-------|-------------------|
1120
- | **Archetype** | 4 | Grounding (hallucination resistance), Substitution (scope adherence), Distractor (context noise), Recovery (failure handling) |
1121
- | **Performance** | 3 | Latency, token efficiency, parallel execution capability |
1122
- | **Quality** | 3 | Output format compliance, correct tool usage, scope adherence |
1415
+ > Evaluate the release-note drafting capability against this repository. Use only committed changelog entries and
1416
+ merged PR metadata. The result is acceptable if every claim links to a source artifact and uncited claims are listed
1417
+ separately.
1123
1418
 
1124
- Target score: >=85% per agent. Results include passed/failed breakdown with evidence.
1125
-
1126
- ---
1419
+ That prompt-led path is intentional. AIWG can route the work through provider-native tools, but the acceptance
1420
+ criterion remains explicit and reviewable. Avoid treating any score, pass rate, or runtime as guaranteed across models
1421
+ or providers; those values depend on the selected model, available tools, project size, and the evidence the workflow
1422
+ can inspect.
1127
1423
 
1128
1424
  ## Bidirectional Traceability — @-Mention System
1129
1425
 
1130
- Link requirements to architecture to code to tests with semantic @-mentions:
1426
+ AIWG uses lightweight `@` references to connect instructions, generated artifacts, source files, evidence, and
1427
+ follow-up work. The goal is traceability across providers: a model can move from a rule to the code it governs, from a
1428
+ generated report to the source data behind it, or from an issue to the artifact that closed it.
1131
1429
 
1132
1430
  ```markdown
1133
- <!-- In a use case document -->
1134
- This use case @implements(UC-001) the authentication flow
1135
- described in @architecture(SAD-section-3.2).
1136
-
1137
- <!-- In a test file -->
1138
- // @tests(UC-001) @depends(auth-service)
1139
- describe('authentication flow', () => { ... });
1431
+ <!-- In an agent or skill file -->
1432
+ @src/auth/middleware.ts
1433
+ @docs/security/authentication.md
1434
+ @.aiwg/evidence/release-2026-09-07.json
1140
1435
  ```
1141
1436
 
1142
- **Mention Commands:**
1437
+ Traceability matters most when a workflow crosses boundaries. An SDLC intake can reference the use case it created. A
1438
+ context-firewall review can reference the baseline it approved. A dataset query can reference the source, ingest plan,
1439
+ checkpoint, and verification record instead of relying on conversation memory.
1440
+
1441
+ Provider-facing commands and skills may expose mention helpers such as mention wiring, validation, linting, or
1442
+ reporting. Because providers package commands differently, the portable entry point is discovery:
1143
1443
 
1144
1444
  ```bash
1145
- /mention-wire # Analyze codebase and inject @-mentions for traceability
1146
- /mention-validate # Validate all @-mentions resolve to existing files
1147
- /mention-report # Generate traceability report
1148
- /mention-lint # Lint @-mentions for style consistency
1149
- /mention-conventions # Display naming conventions and placement rules
1445
+ aiwg discover "mention validate" --limit 5
1446
+ aiwg discover "traceability report" --limit 5
1447
+ aiwg show skill context-firewall
1150
1448
  ```
1151
1449
 
1152
- Relationship qualifiers: `@implements`, `@tests`, `@depends`, `@derives-from`, `@blocked-by`, `@supersedes`. Enables queries like "what implements UC-001?" and "what tests cover the auth module?"
1450
+ When deployed into a provider that supports slash commands, the same work is often available as a prompt command, for example:
1153
1451
 
1154
- ---
1452
+ ```text
1453
+ /mention-validate --target docs
1454
+ /mention-report --scope .aiwg/reports
1455
+ ```
1456
+
1457
+ Use root-relative references in public documentation so links work from the README. Use provider-specific absolute
1458
+ paths only inside generated provider files where that provider requires them.
1155
1459
 
1156
1460
  ## Configuration & Customization
1157
1461
 
1158
- ### Workspace Structure
1159
-
1160
- ```
1161
- your-project/
1162
- ├── .aiwg/ # SDLC artifacts (persistent project memory)
1163
- │ ├── intake/ # Project intake forms
1164
- │ ├── requirements/ # Use cases, user stories, NFRs
1165
- │ ├── architecture/ # SAD, ADRs, system diagrams
1166
- │ ├── planning/ # Phase plans, iteration plans
1167
- │ ├── risks/ # Risk register, mitigations
1168
- │ ├── testing/ # Test strategy, plans, results
1169
- │ ├── security/ # Threat models, security gates
1170
- │ ├── deployment/ # Deployment plans, runbooks
1171
- │ ├── reports/ # Generated status reports
1172
- │ ├── ralph/ # In-session agent loop state
1173
- │ ├── ralph-external/ # External Ralph crash-resilient state
1174
- │ ├── research/ # Research corpus and findings
1175
- │ ├── forensics/ # Investigation artifacts
1176
- │ ├── working/ # Temporary files (safe to delete)
1177
- │ └── frameworks/ # Installed framework registry
1178
- │ └── registry.json
1179
- ├── .claude/ # Claude Code deployment
1180
- │ ├── agents/ # 162 agent definitions
1181
- │ ├── commands/ # Slash commands
1182
- │ ├── skills/ # 86 skill definitions
1183
- │ └── rules/ # RULES-INDEX.md + on-demand full rules
1184
- ├── .github/ # GitHub Copilot deployment
1185
- ├── .cursor/ # Cursor deployment
1186
- ├── .warp/ # Warp Terminal deployment
1187
- └── CLAUDE.md # Project instructions (auto-generated)
1462
+ AIWG separates project context from provider packaging. The project keeps canonical context and generated artifacts
1463
+ under the workspace, then deploys provider-specific adapters for Claude, Codex, Cursor, Windsurf, Warp, OpenCode,
1464
+ OpenClaw, OpenHuman, Hermes, DeepSeek Harness, Copilot, Devin, Factory, Oh My Pi, Pi Coding Agent, Antigravity, and
1465
+ the generic fallback.
1466
+
1467
+ The primary files are:
1468
+
1469
+ ```text
1470
+ WORKSPACE.md # Project/operator context read by providers
1471
+ AIWG.md # AIWG discovery and routing guide
1472
+ AGENTS.md # Provider bootstrap for Codex and other AGENTS.md readers
1473
+ .aiwg/ # Canonical AIWG config, generated context, evidence, reports
1474
+ .aiwg/aiwg.config # Workspace configuration and provider deployment state
1475
+ .aiwg/index/ # Searchable artifact and capability indexes when generated
1476
+ .aiwg/reports/ # Audits, sync reports, doctor reports, workflow outputs
1477
+ .aiwg/sessions/ # Optional local session catalog data
1478
+ .aiwg/datasets/ # Optional dataset plans, manifests, lineage, and exports
1479
+ ```
1480
+
1481
+ Provider directories are generated from the same canonical context. Their exact shape depends on the provider:
1482
+
1483
+ ```text
1484
+ .claude/ # Claude Code skills, commands, hooks, settings
1485
+ .codex/ or ~/.codex/ # Codex prompts and global configuration where applicable
1486
+ .agents/ # Cross-provider agents and skills used by Codex/Antigravity/OMP
1487
+ .cursor/ # Cursor rules and skills
1488
+ .github/ # GitHub Copilot prompts, instructions, and agents
1489
+ .warp/ # Warp skills and compatibility assets
1490
+ .omp/ # Oh My Pi native agents, prompts, rules, and bootstrap
1188
1491
  ```
1189
1492
 
1190
1493
  ### Creating Custom Extensions
1191
1494
 
1495
+ Use the current scaffolding commands for new AIWG assets. The older scaffold commands remain available in some
1496
+ workspaces for compatibility, but the `new-*` and `add-*` commands are the clearer path for new work.
1497
+
1498
+ ```bash
1499
+ # Create a new bundle, extension, addon, framework, or provider adapter
1500
+ aiwg new-bundle my-bundle
1501
+ aiwg new-extension my-extension
1502
+ aiwg new-addon my-addon
1503
+ aiwg new-framework my-framework
1504
+ aiwg new-provider my-provider
1505
+
1506
+ # Add individual provider-facing assets
1507
+ aiwg add-agent release-reviewer --framework sdlc
1508
+ aiwg add-command release-checklist --framework sdlc
1509
+ aiwg add-skill release-notes --framework sdlc
1510
+
1511
+ # Validate metadata before sharing or deploying
1512
+ aiwg validate-metadata
1513
+ ```
1514
+
1515
+ A custom extension should define the smallest durable contract needed by the workflow: triggers, inputs, outputs,
1516
+ evidence, and provider packaging. Keep model-specific phrasing in provider assets. Keep project policy, schemas, and
1517
+ reusable workflow contracts in `.aiwg` or extension source so multiple providers can share them.
1518
+
1519
+ ### Capability Discovery — `aiwg discover` + `aiwg show`
1520
+
1521
+ `aiwg discover` searches the installed AIWG capability catalog. It is the recommended entry point when you know the
1522
+ task but not the command, skill, agent, or framework name.
1523
+
1192
1524
  ```bash
1193
- # Add a custom agent
1194
- aiwg add-agent my-domain-expert
1525
+ # Find capabilities by plain-language intent
1526
+ aiwg discover "deploy production" --limit 5
1527
+ aiwg discover "dataset lineage" --type skill --limit 5
1528
+ aiwg discover "SDLC intake requirements" --limit 8
1529
+
1530
+ # Inspect the selected capability before running or asking a provider to use it
1531
+ aiwg show skill aiwg-status
1532
+ aiwg show skill dataset-intelligence
1533
+ aiwg show skill rlm-prep
1534
+ ```
1195
1535
 
1196
- # Add a custom command
1197
- aiwg add-command my-workflow
1536
+ Discovery is also useful for documentation. Instead of hard-coding every command in a README, link to the relevant
1537
+ quickstart and show one or two representative commands. The catalog can change as frameworks and addons are installed,
1538
+ while the task language stays stable.
1198
1539
 
1199
- # Add a custom skill
1200
- aiwg add-skill my-capability
1540
+ ### Artifact Index — `aiwg index`
1201
1541
 
1202
- # Scaffold a complete addon
1203
- aiwg scaffold-addon my-addon
1542
+ The artifact index makes generated work easier to find, verify, and reuse. It indexes reports, generated context,
1543
+ evidence, and other AIWG-managed files into a searchable local catalog.
1204
1544
 
1205
- # Scaffold a complete framework
1206
- aiwg scaffold-framework my-framework
1545
+ ```bash
1546
+ # Build or refresh the local artifact index
1547
+ aiwg index
1207
1548
 
1208
- # Validate all extension metadata
1209
- aiwg validate-metadata
1549
+ # Inspect status after deployment or refresh
1550
+ aiwg status --probe
1551
+ aiwg doctor
1210
1552
  ```
1211
1553
 
1212
- ### Capability Discovery — `aiwg discover` + `aiwg show`
1554
+ A provider can then answer questions such as “find the latest context-firewall report” or “show the SDLC artifact that
1555
+ introduced this acceptance criterion” without scanning the whole repository manually.
1556
+
1557
+ ### Doc Sync — Bidirectional Documentation
1213
1558
 
1214
- The headline operator surface for finding and reading AIWG capabilities. Most AIWG skills (~455 of 480+) are **not loaded into your platform's flat skill listing** — they stay at `$AIWG_ROOT` and are reached on demand through `aiwg discover` (find) and `aiwg show` (fetch). The kernel set on disk is small on purpose: 9 framework quickrefs + 16 self-maintenance and discovery skills = 25 skills, within supported provider listing budgets.
1559
+ Doc sync is for keeping code and documentation aligned under review. It can audit mismatches, propose updates, and
1560
+ write reports before code or documentation changes are accepted.
1215
1561
 
1216
1562
  ```bash
1217
- # Find a skill by capability
1218
- aiwg discover "deploy production" # → flow-deploy-to-production
1219
- aiwg discover "create intake" # → intake-* family
1220
- aiwg discover "audit security" --type skill --limit 5
1221
- aiwg discover "<phrase>" --format json # stable ids for sub-agents, no paths
1222
- aiwg discover "<phrase>" --format json --compact
1563
+ # Audit both directions without writing changes
1564
+ aiwg doc-sync full --dry-run --scope docs
1223
1565
 
1224
- # Fetch the full body of a specific artifact (companion to discover)
1225
- aiwg show skill aiwg:skill:6f1477d99813ca8d
1226
- aiwg show skill flow-deploy-to-production
1227
- aiwg show agent aiwg-steward
1228
- aiwg show command discover
1229
- aiwg show rule no-attribution
1566
+ # Propose documentation updates from code changes
1567
+ aiwg doc-sync code-to-docs --scope src --guidance "Update quickstarts only"
1230
1568
 
1231
- # Inspect Fortemi metadata and resolved paths when needed
1232
- aiwg show metadata aiwg:skill:6f1477d99813ca8d --json
1569
+ # Propose code TODOs or implementation tasks from documentation requirements
1570
+ aiwg doc-sync docs-to-code --scope docs --interactive
1233
1571
  ```
1234
1572
 
1235
- The kernel quickrefs ship **curated, validated discovery phrases per capability domain** — phrases tested against the live scorer to surface the right top-3 candidates. The self-maintenance and discovery set (including `steward`, `aiwg-doctor`, `aiwg-refresh`, `aiwg-status`, `aiwg-help`, and `use`) stays loaded so the agent retains repair surfaces even when discovery itself is broken. See [`docs/discovery-and-kernel-skills.md`](docs/discovery-and-kernel-skills.md) for the full best-practices guide, ASCII flow diagrams, and verification steps.
1573
+ Doc sync writes reports under `.aiwg/working/` and `.aiwg/reports/` when configured. Treat those reports as review
1574
+ artifacts. Do not assume doc sync can prove semantic equivalence between code and prose; it identifies
1575
+ inconsistencies, stale examples, missing links, and candidate updates for human or provider review.
1236
1576
 
1237
- ### Artifact Index — `aiwg index`
1577
+ ### Reproducibility Validation
1578
+
1579
+ AIWG’s reproducibility features focus on explicit inputs, evidence records, deterministic modes where available, and
1580
+ reviewable outputs. They do not guarantee identical model text across providers or runs.
1238
1581
 
1239
1582
  ```bash
1240
- # Build searchable artifact index
1241
- aiwg index build
1242
- aiwg index build --force --verbose
1583
+ # Put local execution in a stricter mode for workflows that honor it
1584
+ aiwg execution-mode strict --seed 12345
1243
1585
 
1244
- # Search artifacts by keyword
1245
- aiwg index query "authentication" --json
1586
+ # Export and verify evidence when workflows emit evidence bundles
1587
+ aiwg evidence export --help
1588
+ aiwg evidence verify --help
1246
1589
 
1247
- # Show dependency graph for an artifact
1248
- aiwg index deps .aiwg/requirements/UC-001.md --json
1590
+ # Verify workspace health and generated provider context
1591
+ aiwg verify --help
1592
+ aiwg doctor
1593
+ ```
1249
1594
 
1250
- # Index statistics
1251
- aiwg index stats --json
1595
+ For workflows that expose checkpointing, snapshots, or replay through installed skills, start with discovery so the
1596
+ current workspace selects the correct implementation:
1597
+
1598
+ ```bash
1599
+ aiwg discover "create checkpoint" --limit 5
1600
+ aiwg discover "replay evidence" --limit 5
1252
1601
  ```
1253
1602
 
1254
- The index supports multiple graphs: project graph (`.aiwg/` artifacts), codebase graph (`src/` / `test/` / `tools/`), and framework graph (`agentic/code/` + `docs/`).
1603
+ The practical standard is repeatability of inputs, citations, commands, and artifacts. Exact model wording should be
1604
+ treated as a generated output, not as the source of truth.
1255
1605
 
1256
- ### Doc Sync — Bidirectional Documentation
1606
+ ### Session Catalog
1607
+
1608
+ The session catalog is an optional local feature for importing, searching, and promoting useful provider conversation
1609
+ history. It is designed for controlled handoff and audit. It should be enabled intentionally because it can include
1610
+ sensitive prompts, local paths, and project context.
1257
1611
 
1258
1612
  ```bash
1259
- # Audit doc drift (dry run)
1260
- aiwg doc-sync code-to-docs --dry-run
1613
+ # Install the SQLite feature before using the session catalog
1614
+ aiwg features install sqlite
1261
1615
 
1262
- # Sync docs to match code
1263
- aiwg doc-sync code-to-docs
1616
+ # Discover importable sessions for this workspace without changing state
1617
+ aiwg sessions discover --workspace "$PWD" --dry-run
1264
1618
 
1265
- # Bidirectional reconciliation
1266
- aiwg doc-sync full --interactive
1619
+ # Import discovered sessions after review
1620
+ aiwg sessions import-discovered --workspace "$PWD" --confirm
1621
+
1622
+ # Inspect, search, and audit imported sessions
1623
+ aiwg sessions list
1624
+ aiwg sessions timeline
1625
+ aiwg sessions search "release blocker"
1626
+ aiwg sessions doctor
1267
1627
  ```
1268
1628
 
1269
- ### Reproducibility Validation
1629
+ Imported sessions can be tagged, extracted into reusable notes, reviewed for promotion, or audited for provenance.
1630
+ Keep private-provider roots and shared history locations explicit in configuration; do not assume another provider’s
1631
+ global history is safe to import by default.
1632
+
1633
+ See [session history setup](docs/getting-started/session-history.md) and [sessions CLI](docs/sessions/cli.md).
1634
+
1635
+ ### Dataset Intelligence
1636
+
1637
+ Dataset intelligence gives AIWG a governed path for local files, directories, CSV/JSONL sources, and approved HTTP
1638
+ sources. The dataset router carries stable source, plan, checkpoint, lineage, verification, and export references
1639
+ between phases.
1270
1640
 
1271
1641
  ```bash
1272
- # Show/set execution mode (strict = temperature 0, fixed seed)
1273
- aiwg execution-mode
1642
+ # Register a source from a JSON descriptor
1643
+ aiwg dataset source --file dataset-source.json --json
1274
1644
 
1275
- # Create execution snapshot
1276
- aiwg snapshot
1645
+ # Check and preview before ingestion
1646
+ aiwg dataset check source:docs --json
1647
+ aiwg dataset preview source:docs --count 5 --offline
1277
1648
 
1278
- # Create workflow checkpoint
1279
- aiwg checkpoint
1649
+ # Create and approve an ingest plan
1650
+ aiwg dataset plan --file dataset-plan.json --json
1651
+ aiwg dataset ingest plan:docs-index \
1652
+ --digest sha256:<approved-plan-digest> \
1653
+ --idempotency-key docs-index-2026-09-07
1280
1654
 
1281
- # Validate workflow reproducibility
1282
- aiwg reproducibility-validate
1655
+ # Inspect and use the resulting dataset
1656
+ aiwg dataset status dataset:docs-index
1657
+ aiwg dataset verify dataset:docs-index
1658
+ aiwg dataset query dataset:docs-index "Which quickstart explains Codex setup?"
1659
+ aiwg dataset lineage dataset:docs-index
1660
+ aiwg dataset export dataset:docs-index --json
1283
1661
  ```
1284
1662
 
1285
- Thresholds: compliance audit (100%), security scan (100%), test generation (95%).
1663
+ Local adapters are constrained by configured roots. HTTP adapters are deny-by-default and require explicit hosts.
1664
+ Indexes are derived artifacts; the canonical record is the source descriptor, approved plan, ingest run, and evidence
1665
+ trail.
1286
1666
 
1287
- ---
1667
+ See [dataset intelligence quickstart](docs/addons/dataset-intelligence/quickstart.md), [dataset
1668
+ overview](docs/addons/dataset-intelligence/overview.md), and [source adapters](docs/dataset/source-adapters.md).
1288
1669
 
1289
1670
  ## Issue-Driven Development
1290
1671
 
1291
- AIWG integrates with issue trackers for 2-way human-AI collaboration:
1672
+ AIWG supports local issue planning and governed handoff to external issue trackers. The local issue CLI stores records
1673
+ under `.aiwg/issues/`, which makes issues reviewable even when a project does not have GitHub, Gitea, Jira, or another
1674
+ tracker connected.
1292
1675
 
1293
1676
  ```bash
1294
- # Create issues from any backend (Gitea, GitHub, Jira, Linear, local files)
1295
- /issue-create "Implement OAuth2 flow" --labels "feature,auth"
1677
+ # Initialize a local issue store for this workspace
1678
+ aiwg issue init --prefix APP
1679
+
1680
+ # Draft a new local issue
1681
+ aiwg issue plan \
1682
+ --title "Implement OAuth2 callback validation" \
1683
+ --body "Add state validation, token exchange error handling, and tests."
1296
1684
 
1297
- # List and filter issues
1298
- /issue-list --state open --labels "priority:high"
1685
+ # Review and update issues locally
1686
+ aiwg issue list --status open --label auth --limit 20
1687
+ aiwg issue show APP-0001 --comments last:10
1688
+ aiwg issue comment APP-0001 --body "Validated the callback edge cases."
1689
+ aiwg issue close APP-0001 --reason "Implemented and tested."
1690
+ ```
1299
1691
 
1300
- # Drive an issue with agent loop — posts status to issue thread
1301
- /issue-driven-ralph 42
1692
+ External tracker import/export is explicit. Use it when you need traceability between local AIWG records and a remote
1693
+ system, and keep snapshots or live connector settings under review.
1302
1694
 
1303
- # Auto-sync issues from commits and artifacts
1304
- /issue-sync
1695
+ ```bash
1696
+ # Import a tracker snapshot into the local issue store
1697
+ aiwg issue import --from github --snapshot-file issues-snapshot.json
1305
1698
 
1306
- # Close with comprehensive summary and verification
1307
- /issue-close 42
1699
+ # Export a local issue payload for a tracker
1700
+ aiwg issue export APP-0001 --to gitea --out APP-0001.gitea.json
1701
+
1702
+ # Inspect conflicts when reconciling local and external state
1703
+ aiwg issue sync conflicts APP-0001 --snapshot-file issue-APP-0001.json
1308
1704
  ```
1309
1705
 
1310
- The `/address-issues` command orchestrates issue-thread-driven agent loops with automatic progress posting and human feedback incorporation at each cycle.
1706
+ For agent-assisted repair work, ask for the issue outcome directly and include the acceptance checks. In providers
1707
+ with deployed prompt commands, `/address-issues` can route the work through the configured workflow.
1311
1708
 
1312
- ---
1709
+ ```text
1710
+ /address-issues APP-0001 APP-0002 --checks "npm test && npm run lint"
1711
+ ```
1712
+
1713
+ External issue systems are not a default side effect of `aiwg issue`. They require configured connectors, snapshots,
1714
+ or explicit export/import commands. See [local issue integration](docs/local-issues.md) and [filing
1715
+ issues](docs/contributing/filing-issues.md).
1313
1716
 
1314
1717
  ## Daemon Mode & Messaging Integration
1315
1718
 
1316
- ### Daemon Mode
1719
+ AIWG’s automation layer is for long-running coordination, not for hiding work from review. The safe default is local,
1720
+ explicit execution with visible status and evidence. Daemon, messaging, and mission-control setups should declare
1721
+ their trigger source, operator identity, workspace, budget limits, and completion criteria.
1722
+
1723
+ The base CLI exposes current orchestration commands through Ralph and mission control. Messaging bridges and chat bots
1724
+ are advanced deployments described in the daemon and messaging docs; they require external service configuration and
1725
+ should not be assumed to exist in a fresh checkout.
1317
1726
 
1318
1727
  ```bash
1319
- # Background file watching, cron scheduling, IPC
1320
- aiwg daemon start
1728
+ # Start a managed mission-control session
1729
+ aiwg mc start --name "release follow-up" --max-missions 3
1730
+
1731
+ # Dispatch bounded work with an explicit completion criterion
1732
+ aiwg mc dispatch <session-id> \
1733
+ "Fix the failing auth tests" \
1734
+ --completion "npm test -- auth passes"
1735
+
1736
+ # Inspect and control running work
1737
+ aiwg mc run <session-id>
1738
+ aiwg mc status <session-id>
1739
+ aiwg mc watch <session-id>
1740
+ aiwg mc pause <session-id>
1741
+ aiwg mc resume <session-id>
1742
+ aiwg mc stop <session-id>
1321
1743
  ```
1322
1744
 
1323
- See [Daemon Guide](docs/daemon-guide.md) for background agent orchestration.
1745
+ For provider messaging, document the concrete external channel and approval boundary. A Slack, Discord, Telegram, or
1746
+ webhook bridge should make it clear who can enqueue work, where logs are stored, and which operations require human
1747
+ approval before writing to external systems.
1748
+
1749
+ See [daemon guide](docs/daemon-guide.md), [messaging guide](docs/messaging-guide.md), and [Mission Control](docs/addons/ralph/quickstart.md).
1750
+
1751
+ ## See It In Action
1324
1752
 
1325
- ### Messaging Integration
1753
+ The fastest way to use AIWG is to ask for the first useful task, request a concrete deliverable, and name the success
1754
+ check. Commands help when you know the exact workflow; plain-language task prompts are better when AIWG should choose
1755
+ the relevant skill or provider surface.
1326
1756
 
1327
- Bidirectional Slack, Discord, and Telegram bots for remote agent control:
1757
+ ### SDLC workflow from idea to implementation
1328
1758
 
1329
1759
  ```bash
1330
- # Connect to messaging platforms
1331
- aiwg messaging connect slack
1332
- aiwg messaging connect discord
1333
- aiwg messaging connect telegram
1760
+ # Discover the right SDLC entry point
1761
+ aiwg discover "SDLC intake requirements architecture" --limit 5
1762
+
1763
+ # Run the accelerator when you want AIWG to scaffold the SDLC work plan
1764
+ aiwg sdlc-accelerate "AI-powered code review tool" \
1765
+ --success "requirements, architecture notes, and first implementation task are generated"
1334
1766
  ```
1335
1767
 
1336
- See [Messaging Guide](docs/messaging-guide.md) for setup and configuration.
1768
+ Provider prompt:
1337
1769
 
1338
- ---
1770
+ ```text
1771
+ Use the SDLC framework to turn “AI-powered code review tool” into requirements, architecture decisions, a first implementation task, and acceptance checks. Stop with links to the generated artifacts.
1772
+ ```
1339
1773
 
1340
- ## See It In Action
1774
+ ### Long-running implementation loop
1341
1775
 
1342
1776
  ```bash
1343
- # Generate project intake from natural language
1344
- /intake-wizard "Build customer portal with real-time chat"
1777
+ aiwg ralph "Fix all failing tests in the auth package" \
1778
+ --completion "npm test -- auth passes" \
1779
+ --max-iterations 5 \
1780
+ --max-wall-clock-minutes 45
1345
1781
 
1346
- # Accelerate from idea to construction-ready
1347
- /sdlc-accelerate "AI-powered code review tool"
1782
+ aiwg ralph-status
1783
+ aiwg ralph-resume <loop-id>
1784
+ aiwg ralph-abort <loop-id>
1785
+ ```
1348
1786
 
1349
- # Phase transition with automated gate check
1350
- /flow-inception-to-elaboration
1787
+ Ralph is useful for bounded repair loops where the success condition is objective. It is not a guarantee that the
1788
+ model will solve the task. Set wall-clock, token, tool-call, or cost limits for expensive providers.
1351
1789
 
1352
- # Iterative task execution — "iteration beats perfection"
1353
- /ralph "Fix all failing tests" --completion "npm test passes"
1790
+ ### Recursive search over large code or docs
1354
1791
 
1355
- # Long-running tasks with crash recovery (6-8 hours)
1356
- /ralph-external "Migrate to TypeScript" --completion "npx tsc --noEmit exits 0"
1792
+ ```bash
1793
+ aiwg rlm-prep docs/ --strategy semantic-boundary --size 200
1794
+ aiwg rlm-search "Where do provider quickstarts mention reload requirements?" \
1795
+ --source .aiwg/rlm-prep/<manifest-dir>/manifest.json \
1796
+ --max-parallel 4 \
1797
+ --budget 50000
1798
+ aiwg rlm-cache stats
1799
+ ```
1357
1800
 
1358
- # Process massive codebases with recursive context decomposition
1359
- /rlm-query "src/**/*.ts" "Extract all exported interfaces" --model haiku
1360
- /rlm-batch "src/components/*.tsx" "Add TypeScript types" --max-parallel 4
1801
+ Provider prompt:
1361
1802
 
1362
- # Digital forensics investigation
1363
- /forensics-investigate
1364
- /forensics-triage
1365
- /forensics-timeline
1803
+ ```text
1804
+ Find every user-facing quickstart that still tells users to manually reload after `aiwg use all`. Return file links, the quoted sentence, and the replacement language.
1805
+ ```
1366
1806
 
1367
- # Scan codebase for agent-readiness
1368
- /codebase-health --format text
1807
+ ### Dataset-backed project knowledge
1369
1808
 
1370
- # Decompose large files into agent-friendly modules
1371
- /decompose-file src/large-file.ts --execute
1809
+ ```bash
1810
+ aiwg dataset check source:docs --json
1811
+ aiwg dataset preview source:docs --count 5 --offline
1812
+ aiwg dataset query dataset:docs-index "Which provider quickstart is best for Codex?"
1813
+ ```
1372
1814
 
1373
- # Deploy to production with rollback gates
1374
- /flow-deploy-to-production
1815
+ Use dataset intelligence when the source and lineage need to be explicit. Use RLM when the immediate need is recursive
1816
+ search or fanout over files.
1375
1817
 
1376
- # Security assessment
1377
- /security-audit
1818
+ ### Session history reuse
1378
1819
 
1379
- # Voice transformation
1380
- "Apply technical-authority voice to docs/architecture.md"
1381
- "Create a voice profile based on our existing blog posts"
1820
+ ```bash
1821
+ aiwg sessions discover --workspace "$PWD" --dry-run
1822
+ aiwg sessions search "marketing audit"
1823
+ aiwg sessions extract <session-id> --format markdown
1382
1824
  ```
1383
1825
 
1384
- ---
1826
+ This is useful when prior provider conversations contain decisions that should become project artifacts. Keep import
1827
+ scope explicit and review the discovered sessions before promotion.
1385
1828
 
1386
- ## Platform Support
1829
+ ### Security, forensics, and operations prompts
1387
1830
 
1388
- AIWG supports 14 named provider integrations. Artifact support varies by provider and is adapted to each provider's native or compatibility surfaces.
1389
-
1390
- | Platform | Status | Agents | Commands | Skills | Rules | Deploy Command |
1391
- |----------|--------|--------|----------|--------|-------|---------------|
1392
- | **[Google Antigravity CLI](docs/providers/antigravity.md)** | Experimental | degraded `.agents/agents/` | — | `.agents/skills/` | `AGENTS.md` | `--provider antigravity` (alias: `agy`) |
1393
- | **Claude Code** | Tested | `.claude/agents/` | `.claude/commands/` | `.claude/skills/` | `.claude/rules/` | `aiwg use sdlc` |
1394
- | **GitHub Copilot** | Tested | `.github/agents/` | `.github/agents/` | `.github/skills/` | `.github/copilot-rules/` | `--provider copilot` |
1395
- | **Warp Terminal** | Tested | `.warp/agents/` + WARP.md | `.warp/commands/` | `.warp/skills/` | `.warp/rules/` | `--provider warp` |
1396
- | **Factory AI** | Tested | `.factory/droids/` | `.factory/commands/` | `.factory/skills/` | `.factory/rules/` | `--provider factory` |
1397
- | **Cursor** | Tested | `.cursor/agents/` | `.cursor/commands/` | `.cursor/skills/` | `.cursor/rules/` | `--provider cursor` |
1398
- | **OpenCode** | Tested | `.opencode/agent/` | `.opencode/commands/` | `.opencode/skill/` | `.opencode/rule/` | `--provider opencode` |
1399
- | **OpenAI/Codex** | Tested | `.codex/agents/` | `~/.codex/prompts/` | `.agents/skills/` | `.codex/rules/` | `--provider codex` |
1400
- | **Devin Desktop** | Tested compatibility adapter | AGENTS.md | `.windsurf/workflows/` | `.windsurf/skills/` | `.windsurf/rules/` | `--provider devin` |
1401
- | **Hermes** | Stable | — | — | `~/.hermes/skills/.aiwg/` | — | `--provider hermes` |
1402
- | **OpenClaw** | Tested | `~/.openclaw/agents/` | `~/.openclaw/commands/` | `~/.openclaw/.aiwg/skills/` | `~/.openclaw/rules/` | `--provider openclaw` |
1403
- | **OpenHuman** | Experimental | — | — | `~/.openhuman/.aiwg/skills/` | `~/.openhuman/.aiwg/rules/` | `--provider openhuman` |
1404
- | **[Pi Coding Agent](https://pi.dev/)** | Experimental | `.agents/skills/` | `.pi/prompts/` | `.agents/skills/` | `AGENTS.md` | `--provider pi` |
1405
- | **[Oh My Pi](docs/providers/omp.md)** | Experimental | `.omp/agents/` | `.omp/prompts/` | `.agents/skills/` | `.omp/AGENTS.md` | `--provider omp` |
1406
-
1407
- The legacy `--provider windsurf` selector remains supported and writes the
1408
- same `.windsurf/` compatibility paths, but new commands should use `devin`.
1409
- `devin-cli` is a distinct product surface and is not currently a deployable
1410
- AIWG provider.
1831
+ ```text
1832
+ Use the security-engineering framework to review the OAuth2 callback flow. Produce threat assumptions, concrete findings, and tests I can run.
1411
1833
 
1412
- ---
1834
+ Use the forensics framework to analyze these logs. Preserve evidence references, build a timeline, and separate confirmed facts from hypotheses.
1413
1835
 
1414
- ## CLI Reference (50 Commands)
1415
-
1416
- | Category | Commands | Description |
1417
- |----------|----------|-------------|
1418
- | **Maintenance** | `help`, `version`, `doctor`, `context-firewall`, `update` | Installation health, context safety, updates, diagnostics |
1419
- | **Framework** | `use`, `list`, `remove` | Deploy, inspect, and remove frameworks |
1420
- | **Project** | `new` | Scaffold new project with AIWG structure |
1421
- | **Workspace** | `status`, `migrate-workspace`, `rollback-workspace` | Workspace health and migration |
1422
- | **MCP** | `mcp serve`, `mcp install`, `mcp info` | Model Context Protocol server |
1423
- | **Catalog** | `catalog list`, `catalog info`, `catalog search` | Browse available extensions |
1424
- | **Marketplace packaging** | `install-plugin`, `uninstall-plugin`, `plugin-status`, `package-plugin`, `package-all-plugins` | Install and package delivery wrappers |
1425
- | **Scaffolding** | `add-agent`, `add-command`, `add-skill`, `add-template`, `scaffold-addon`, `scaffold-extension`, `scaffold-framework` | Create new extensions |
1426
- | **Ralph** | `ralph`, `ralph-status`, `ralph-abort`, `ralph-resume`, `ralph-external`, `ralph-memory`, `ralph-config` | Iterative execution engine |
1427
- | **Metrics & evidence** | `cost-report`, `cost-history`, `metrics-tokens`, `evidence` | Token usage, cost tracking, and portable evaluation evidence |
1428
- | **Index** | `index build`, `index query`, `index deps`, `index stats` | Artifact discovery and dependency graphing |
1429
- | **Documentation** | `doc-sync` | Bidirectional doc-code synchronization |
1430
- | **SDLC** | `sdlc-accelerate` | Idea-to-construction-ready pipeline |
1431
- | **Code Analysis** | `cleanup-audit` | Dead code and unused export detection |
1432
- | **Reproducibility** | `execution-mode`, `snapshot`, `checkpoint`, `reproducibility-validate` | Deterministic workflow validation |
1433
- | **Toolsmith** | `runtime-info` | Runtime environment detection |
1434
- | **Utility** | `prefill-cards`, `contribute-start`, `validate-metadata` | Development utilities |
1435
-
1436
- ### Quick Reference
1836
+ Use the ops framework to turn this production incident into a runbook update, verification checklist, and follow-up issues.
1837
+ ```
1838
+
1839
+ These prompts preserve the original README’s hands-on style while keeping provider behavior accurate: AIWG selects
1840
+ framework capabilities through discovery and provider deployment, and the deliverable remains explicit.
1841
+
1842
+ ## Platform Support
1843
+
1844
+ AIWG has registry-backed named provider integrations plus a generic fallback for tools that read Markdown context but
1845
+ do not have a dedicated adapter. The integrations share product framing: reusable project context and specialist
1846
+ workflows in the AI tools teams already use. Provider distinctions matter because each host has different native
1847
+ surfaces.
1848
+
1849
+ See [cross-platform overview](docs/integrations/cross-platform-overview.md) for the maintained comparison and setup links.
1850
+
1851
+ | Provider | Setup | Primary context | Native or conventional surfaces | Notes |
1852
+ |---|---|---|---|---|
1853
+ | Claude Code | `aiwg use all --provider claude` | `CLAUDE.md` | Skills, commands, hooks, MCP config | Best fit for rich AIWG provider packaging. See [Claude quickstart](docs/integrations/claude-code-quickstart.md). |
1854
+ | Codex | `aiwg use all --provider codex` | `AGENTS.md` | Global prompts, project skills, MCP config | Uses AGENTS.md bootstrap plus `.agents/skills/`. See [Codex quickstart](docs/integrations/codex-quickstart.md). |
1855
+ | Cursor | `aiwg use all --provider cursor` | Rules and skills | `.cursor/rules/*.mdc`, `.cursor/skills/*/SKILL.md` | Cursor rules are native; some assets remain conventional. See [Cursor quickstart](docs/integrations/cursor-quickstart.md). |
1856
+ | Windsurf | `aiwg use all --provider windsurf` | Windsurf rules | Rules and workflows | Uses Windsurf’s local rule model where available. |
1857
+ | OpenCode | `aiwg use all --provider opencode` | Agent/rule context | Provider-local agents, commands, rules | Good for lightweight terminal workflows. |
1858
+ | Gemini CLI | `aiwg use all --provider gemini` | `GEMINI.md` | Commands and context files | Keeps AIWG guidance in Gemini-readable Markdown. |
1859
+ | Qwen Code | `aiwg use all --provider qwen` | `QWEN.md` | Commands and context files | Similar Markdown-first provider packaging. |
1860
+ | Firebase Studio | `aiwg use all --provider firebase` | Studio context | Rules and generated context | Focused on Firebase Studio workspace guidance. |
1861
+ | GitHub Copilot | `aiwg use all --provider copilot` | `.github` assets | Prompts, instructions, agents, MCP config | Uses `.github/prompts/*.prompt.md`, `.github/instructions/*.instructions.md`, and `.github/agents/*.agent.md`. |
1862
+ | Devin | `aiwg use all --provider devin` | Devin-compatible context | Compatibility packaging | Uses compatibility paths where Devin can read project instructions. |
1863
+ | Factory | `aiwg use all --provider factory` | Factory context | Agents and commands where supported | Provider behavior depends on the installed Factory environment. |
1864
+ | Oh My Pi | `aiwg use all --provider omp` | `.omp/AGENTS.md` | Agents, prompts, rules, skills | Dedicated OMP quickstart: [Oh My Pi quickstart](docs/providers/omp.md). |
1865
+ | Pi Coding Agent | `aiwg use all --provider pi` | Pi context | Markdown context and provider adapters | See [Pi quickstart](docs/integrations/pi-quickstart.md). |
1866
+ | Antigravity | `aiwg use all --provider antigravity` | `AGENTS.md` and `.agents/` | Agents, skills, indexed commands, MCP config when enabled | See [Antigravity provider docs](docs/providers/antigravity.md). |
1867
+ | Generic Markdown | `aiwg use all --provider generic` | `AIWG.md` / `WORKSPACE.md` | Markdown instructions | Use when a provider reads repo docs but has no dedicated integration. |
1868
+
1869
+ After any deployment, use status and doctor before relying on the provider context:
1437
1870
 
1438
1871
  ```bash
1439
- # Deploy frameworks
1440
- aiwg use sdlc # SDLC framework
1441
- aiwg use forensics # Forensics framework
1442
- aiwg use all # Everything
1443
- aiwg use sdlc --provider copilot # Deploy to GitHub Copilot
1444
-
1445
- # Project management
1446
- aiwg new my-project # Scaffold new project
1447
- aiwg status # Workspace health
1448
- aiwg doctor # Installation diagnostics
1449
- aiwg context-firewall scan # Provider context, trust, drift, and budget audit
1450
-
1451
- # Iterative execution (Agent Loop)
1452
- aiwg ralph "Fix all tests" --completion "npm test passes"
1453
- aiwg ralph-status # Check loop progress
1454
- aiwg ralph-abort # Cancel running loop
1455
- aiwg ralph-resume # Resume interrupted loop
1456
- aiwg ralph-external "Migrate to TS" --completion "tsc --noEmit exits 0"
1457
-
1458
- # Artifact discovery
1459
- aiwg index build # Build artifact index
1460
- aiwg index query "authentication" --json
1461
- aiwg index deps .aiwg/requirements/UC-001.md --json
1462
-
1463
- # Documentation sync
1464
- aiwg doc-sync code-to-docs --dry-run
1465
- aiwg doc-sync full --interactive
1466
-
1467
- # Metrics
1468
- aiwg cost-report # Agent-native session cost breakdown
1469
- aiwg cost-report --fleet # OpenRouter per-bot MTD spend observation
1470
- aiwg evidence export --output ./evidence # Package evaluation evidence and provenance
1471
- aiwg evidence verify ./evidence # Verify hashes and the bundle checkpoint
1472
- aiwg metrics-tokens # Token usage
1473
-
1474
- # SDLC accelerate
1475
- aiwg sdlc-accelerate "Project description"
1476
- aiwg sdlc-accelerate --from-codebase .
1872
+ aiwg status --probe
1873
+ aiwg doctor
1477
1874
  ```
1478
1875
 
1479
- ---
1876
+ Fresh deployments may report that the provider should be restarted or reloaded so the host notices new files. That is
1877
+ provider-specific readiness information, not a requirement to regenerate context after every command.
1480
1878
 
1481
- ## Architecture
1879
+ ## CLI Reference
1482
1880
 
1483
- ### Extension System
1881
+ The CLI is organized around framework deployment, workspace health, discovery, governed artifacts, orchestration, and
1882
+ specialized addons. Use `aiwg help` for the current top-level surface and [CLI reference](docs/cli/reference.md) for
1883
+ the generated reference.
1484
1884
 
1485
- AIWG uses a unified extension system with 10 extension types, projected onto the supported artifact surfaces of 14 named provider integrations:
1486
-
1487
- | Type | Count | Description |
1488
- |------|-------|-------------|
1489
- | **Agents** | 188 | Specialized AI personas with defined tools, responsibilities, and operating rhythms |
1490
- | **Commands** | 50 | CLI commands and slash commands for workflow automation |
1491
- | **Skills** | 128 | Natural language workflow triggers activated by conversation patterns |
1492
- | **Rules** | 35 | Enforcement patterns deployed as consolidated index with on-demand full-rule loading |
1493
- | **Templates** | 334 | Progressive disclosure document templates for all SDLC phases |
1494
- | **Frameworks** | 8 | Complete workflow systems (SDLC, Forensics, Marketing, Research, Media Curator, Ops, Knowledge Base, Security Engineering) |
1495
- | **Addons** | 21 | Feature bundles extending frameworks (RLM, Voice, Testing Quality, UAT, Ring) |
1496
- | **Hooks** | varies | Lifecycle event handlers (pre-session, post-write, workflow tracing) |
1497
- | **Tools** | varies | External utility integrations (git, jq, npm) |
1498
- | **MCP Servers** | varies | Model Context Protocol server integrations |
1885
+ | Area | Commands | Use when |
1886
+ |---|---|---|
1887
+ | Framework deployment | `aiwg use`, `aiwg list`, `aiwg remove` | Install or remove AIWG framework/provider assets in a workspace. |
1888
+ | Getting started | `aiwg init`, `aiwg setup project`, `aiwg new`, `aiwg quickref generate` | Bootstrap a workspace or generate quick reference docs. |
1889
+ | Workspace health | `aiwg status`, `aiwg doctor`, `aiwg refresh`, `aiwg installation`, `aiwg verify` | Check readiness, repair drift, and validate generated context. |
1890
+ | Catalog and discovery | `aiwg catalog`, `aiwg discover`, `aiwg show`, `aiwg index`, `aiwg artifacts` | Find capabilities and locate generated outputs. |
1891
+ | Provider and MCP | `aiwg mcp serve`, `aiwg mcp install`, `aiwg mcp info`, `aiwg runtime-info` | Connect AIWG to MCP hosts or inspect runtime details. |
1892
+ | Execution and dispatch | `aiwg run skill`, `aiwg run script`, `aiwg output-mode`, `aiwg execution-mode` | Invoke portable skills/scripts and control output or reproducibility mode. |
1893
+ | Ralph loop | `aiwg ralph`, `aiwg ralph-status`, `aiwg ralph-resume`, `aiwg ralph-abort`, `aiwg ralph-attach` | Run bounded iterative implementation loops. |
1894
+ | Mission control | `aiwg mc start`, `aiwg mc dispatch`, `aiwg mc status`, `aiwg mc watch`, `aiwg mc stop` | Coordinate multiple bounded missions from one workspace. |
1895
+ | Sessions | `aiwg sessions discover`, `aiwg sessions import-discovered`, `aiwg sessions list`, `aiwg sessions search`, `aiwg sessions doctor` | Import, inspect, and promote provider session history. |
1896
+ | Dataset intelligence | `aiwg dataset source`, `aiwg dataset check`, `aiwg dataset preview`, `aiwg dataset plan`, `aiwg dataset ingest`, `aiwg dataset verify`, `aiwg dataset query` | Govern source intake, indexing, lineage, and queries. |
1897
+ | Evidence and metrics | `aiwg evidence export`, `aiwg evidence verify`, `aiwg cost-report --fleet` | Preserve verification records and inspect spend or usage where configured. |
1898
+ | Scaffolding | `aiwg new-bundle`, `aiwg new-extension`, `aiwg new-addon`, `aiwg new-framework`, `aiwg new-provider`, `aiwg add-agent`, `aiwg add-command`, `aiwg add-skill` | Create new AIWG packages and provider-facing assets. |
1899
+ | Issues | `aiwg issue init`, `aiwg issue plan`, `aiwg issue list`, `aiwg issue show`, `aiwg issue import`, `aiwg issue export` | Maintain local issue records and exchange snapshots with external trackers. |
1499
1900
 
1500
- ### Multi-Agent Orchestration
1901
+ Common setup and inspection flow:
1501
1902
 
1903
+ ```bash
1904
+ # Install all AIWG assets for the current provider
1905
+ aiwg use all --provider codex
1906
+
1907
+ # Verify generated context and provider readiness
1908
+ aiwg status --probe --json
1909
+ aiwg doctor
1910
+
1911
+ # Find and inspect capabilities instead of guessing command names
1912
+ aiwg discover "release planning" --limit 5
1913
+ aiwg show skill aiwg-status
1502
1914
  ```
1503
- ┌─────────────────────┐
1504
- │ Executive Orchestrator│
1505
- └──────────┬──────────┘
1506
- │
1507
- ┌────────────────┼────────────────┐
1508
- ▼ ▼ ▼
1509
- ┌──────────────┐ ┌──────────────┐ ┌──────────────┐
1510
- │Primary Author│ │ Reviewer 1 │ │ Reviewer 2 │ ← Parallel
1511
- │(e.g. Req. │ │(e.g. Security│ │(e.g. Test │
1512
- │ Analyst) │ │ Architect) │ │ Architect) │
1513
- └──────┬───────┘ └──────┬───────┘ └──────┬───────┘
1514
- └────────────────┼────────────────┘
1515
- ▼
1516
- ┌─────────────────────┐
1517
- │ Documentation │
1518
- │ Synthesizer │ ← Merge all reviews
1519
- └──────────┬──────────┘
1520
- ▼
1521
- ┌─────────────────────┐
1522
- │ Human Gate │ ← GO / NO_GO decision
1523
- └──────────┬──────────┘
1524
- ▼
1525
- ┌─────────────────────┐
1526
- │ .aiwg/ Archive │ ← Persistent artifacts
1527
- └─────────────────────┘
1915
+
1916
+ ## Architecture
1917
+
1918
+ AIWG is a portable context and workflow layer. It keeps canonical project instructions in repo-visible files, packages
1919
+ provider-specific assets for the AI tools a team uses, and preserves artifacts so work can be reviewed outside the
1920
+ original chat.
1921
+
1922
+ ```mermaid
1923
+ flowchart TD
1924
+ A[Project context<br/>WORKSPACE.md + AIWG.md] --> B[AIWG catalog]
1925
+ B --> C[Provider packaging]
1926
+ C --> D[Claude, Codex, Cursor, Copilot, Warp, OMP, Antigravity, others]
1927
+ B --> E[Specialist workflows]
1928
+ E --> F[SDLC, research, ops, security, marketing, datasets, RLM]
1929
+ E --> G[Artifacts and evidence]
1930
+ G --> H[Reports, issues, datasets, sessions, indexes]
1931
+ H --> B
1528
1932
  ```
1529
1933
 
1530
- ### YAML Metalanguage
1934
+ ### Extension System
1531
1935
 
1532
- AIWG is pioneering a declarative YAML metalanguage for multi-agent workflow orchestration. Schema-validated YAML defines agent topology, workflow DAGs, gate conditions, and artifact contracts — while natural language handles behavioral logic.
1936
+ An AIWG extension usually contains some combination of:
1533
1937
 
1534
- ```yaml
1535
- # Example: flow definition (schema-validated)
1536
- flow:
1537
- id: inception-to-elaboration
1538
- model: opus
1539
- entry_criteria:
1540
- gate: LOM
1541
- artifacts:
1542
- - path: .aiwg/requirements/vision-document.md
1543
- required: true
1544
- steps:
1545
- - id: requirements-analysis
1546
- agent: requirements-analyst
1547
- parallel_group: reviews
1548
- - id: architecture-baseline
1549
- agent: architecture-designer
1550
- parallel_group: reviews
1551
- - id: synthesis
1552
- agent: documentation-synthesizer
1553
- depends_on: [requirements-analysis, architecture-baseline]
1554
- exit_criteria:
1555
- gate: ABM
1556
- decision: [GO, CONDITIONAL_GO, NO_GO]
1557
- ```
1558
-
1559
- JSON Schema definitions for `flow.yaml`, `agent.yaml`, `rule.yaml`, and `skill.yaml` at `agentic/code/frameworks/sdlc-complete/schemas/metalanguage/`.
1560
-
1561
- ### Project Artifacts (.aiwg/)
1562
-
1563
- All SDLC artifacts persist in `.aiwg/` — structured project memory that survives across AI sessions:
1564
-
1565
- ```
1566
- .aiwg/
1567
- ├── intake/ # Project intake forms, solution profiles
1568
- ├── requirements/ # Use cases, user stories, NFRs
1569
- ├── architecture/ # SAD, ADRs, diagrams
1570
- ├── planning/ # Phase plans, iteration plans
1571
- ├── risks/ # Risk register, mitigations
1572
- ├── testing/ # Test strategy, test plans
1573
- ├── security/ # Threat models, security gates
1574
- ├── deployment/ # Deployment plans, runbooks
1575
- ├── reports/ # Generated status reports
1576
- ├── ralph/ # Agent loop state and history
1577
- └── frameworks/ # Installed framework registry
1578
- ```
1579
-
1580
- This segmentation is what makes large projects manageable. Individual code files inevitably grow, but the project knowledge stays organized into focused domains. An agent working on a deployment problem loads `@.aiwg/deployment/` and `@.aiwg/architecture/` — not the entire codebase. An agent debugging a test failure loads the relevant requirement, the test plan, and the specific source file. Context stays sharp regardless of project size.
1581
-
1582
- `aiwg index` amplifies this further — it builds a searchable artifact index so agents resolve lookups in a single query instead of browsing. Without tooling: 3-6 documents to find what's needed. With AIWG structure: 2-3. With the index: usually 1.
1938
+ | Asset | Role |
1939
+ |---|---|
1940
+ | Agents | Persistent role definitions, responsibilities, and routing constraints. |
1941
+ | Skills | Task-specific procedures with triggers, inputs, outputs, and evidence rules. |
1942
+ | Commands | Provider-facing shortcuts or prompt templates. |
1943
+ | Rules | Policies and reusable constraints. |
1944
+ | Schemas | Structured contracts for plans, artifacts, manifests, and reports. |
1945
+ | Templates | Repeatable starting points for generated files. |
1946
+ | Scripts | Local deterministic helpers used by workflows. |
1583
1947
 
1584
- ---
1948
+ The registry and discovery index make those assets findable without requiring every provider to support every asset
1949
+ type natively. When a provider lacks a native concept, AIWG packages the asset as Markdown context or a conventional
1950
+ file the provider can read.
1585
1951
 
1586
- ## Agent Loop — Autonomous Long-Running Agent Orchestration
1952
+ ### Multi-Agent Orchestration
1587
1953
 
1588
- The Agent Loop is the core execution philosophy: **iteration beats perfection**. Instead of getting everything right on the first attempt, the agent executes in a retry loop where errors become learning data. Ralph supports both in-session loops and **crash-resilient external loops that run indefinitely** — surviving process crashes, terminal disconnects, and system reboots.
1954
+ AIWG’s orchestration model is explicit about roles and handoffs. A complex task can move through a steward, specialist
1955
+ skill, review step, evidence export, and follow-up issue without losing the artifact trail.
1589
1956
 
1590
- ### In-Session Ralph (Minutes to Hours)
1957
+ ```mermaid
1958
+ sequenceDiagram
1959
+ participant U as User
1960
+ participant S as Steward / Discover
1961
+ participant W as Specialist Workflow
1962
+ participant P as Provider Tooling
1963
+ participant E as Evidence Store
1591
1964
 
1592
- ```bash
1593
- # Iterative task execution with automatic error recovery
1594
- /ralph "Fix all failing tests" --completion "npm test passes with 0 failures"
1595
- /ralph "Reach 80% coverage" --completion "coverage report shows >80%" --max-iterations 20
1965
+ U->>S: Describe outcome and success check
1966
+ S->>W: Select capability and inputs
1967
+ W->>P: Execute provider-local or CLI steps
1968
+ P-->>W: Results, files, diagnostics
1969
+ W->>E: Write report/evidence/issue refs
1970
+ W-->>U: Deliverable with checks and next step
1971
+ ```
1972
+
1973
+ This structure is why AIWG documentation emphasizes “first useful task, concrete deliverable, success check, and next
1974
+ step.” It gives a model enough direction to act while leaving a reviewer enough evidence to verify the result.
1975
+
1976
+ ### YAML Metalanguage
1977
+
1978
+ Many AIWG assets use YAML frontmatter or YAML schemas so capabilities can be discovered, validated, and converted
1979
+ between provider formats.
1596
1980
 
1597
- # Issue-driven Ralph — posts cycle status to issue threads, incorporates human feedback
1598
- /issue-driven-ralph 42 # Drives issue #42 with 2-way human-AI collaboration
1981
+ ```yaml
1982
+ ---
1983
+ namespace: aiwg
1984
+ name: release-notes
1985
+ description: Draft release notes from approved issue and changelog artifacts
1986
+ platforms: [all]
1987
+ triggers:
1988
+ - release notes
1989
+ - changelog summary
1990
+ outputs:
1991
+ - docs/releases/{version}.md
1992
+ evidence:
1993
+ required:
1994
+ - source_issue_refs
1995
+ - changelog_refs
1996
+ ---
1599
1997
  ```
1600
1998
 
1601
- ### External Ralph — Crash-Resilient Autonomous Agents (Hours to Days)
1999
+ The metadata is not decorative. It lets `aiwg discover` find the capability, lets validation detect missing fields,
2000
+ and gives provider adapters enough information to package the asset accurately.
1602
2001
 
1603
- External Ralph runs as a **persistent background process** with PID file tracking, crash recovery, and automatic restart. The agent continues working even if your terminal disconnects or the host reboots.
2002
+ ### Project Artifacts
1604
2003
 
1605
- ```bash
1606
- # Long-running autonomous task (6-8+ hours, survives crashes)
1607
- /ralph-external "Migrate entire codebase to TypeScript" \
1608
- --completion "npx tsc --noEmit exits 0" \
1609
- --timeout 480
2004
+ AIWG-generated artifacts are intentionally ordinary files: Markdown, JSON, YAML, SQLite-backed local stores when
2005
+ enabled, and provider-readable context. That makes them inspectable in a code review and portable across machines.
1610
2006
 
1611
- # Autonomous code review loop
1612
- /ralph-external "Review and fix all security vulnerabilities" \
1613
- --completion "npm audit shows 0 vulnerabilities"
2007
+ Examples include:
1614
2008
 
1615
- # Continuous integration loop
1616
- /ralph-external "Get all tests passing on Node 18 and 22" \
1617
- --completion "npm test passes on both versions"
2009
+ ```text
2010
+ .aiwg/reports/context-firewall-*.md
2011
+ .aiwg/reports/doc-sync-audit-*.md
2012
+ .aiwg/evidence/*.json
2013
+ .aiwg/issues/*.json
2014
+ .aiwg/datasets/**/manifest.json
2015
+ .aiwg/rlm-prep/**/manifest.json
2016
+ .aiwg/sessions/**
1618
2017
  ```
1619
2018
 
1620
- External Ralph features:
2019
+ Artifact indexes reduce manual browsing, but they do not replace source review. Treat indexed results as navigation
2020
+ aids with links back to the original files.
1621
2021
 
1622
- - **Crash resilience** — PID file recovery, automatic restart on process death
1623
- - **Checkpoint system** — saves progress at each iteration boundary, resumes from last checkpoint
1624
- - **Cross-session persistence** — state stored in `.aiwg/ralph-external/`, survives terminal disconnects
1625
- - **Debug memory** — learns from failure patterns across iterations, applies lessons to subsequent attempts
1626
- - **Episodic memory** — `/ralph-reflect` shows accumulated learnings and strategy evolution
1627
- - **Completion reports** — detailed iteration history saved to `.aiwg/ralph/`
2022
+ ## Agent Loop — Autonomous Long-Running Agent Orchestration
1628
2023
 
1629
- ### Scheduled and Remote Agents
2024
+ Ralph is AIWG’s bounded iterative agent loop. It is intended for tasks where the objective and completion criterion
2025
+ can be checked: fixing tests, applying a migration, updating docs to match a report, or carrying a refactor through
2026
+ verification.
1630
2027
 
1631
2028
  ```bash
1632
- # Schedule recurring autonomous agent tasks
1633
- /schedule create "Run security audit" --cron "0 9 * * 1" # Every Monday 9am
1634
- /schedule create "Check dependency updates" --cron "0 0 * *" # Monthly
2029
+ aiwg ralph "Update provider quickstarts from the marketing audit" \
2030
+ --completion "changed files match the approved audit scope and markdown links pass" \
2031
+ --max-iterations 6 \
2032
+ --max-wall-clock-minutes 60 \
2033
+ --max-tool-calls 120
2034
+ ```
1635
2035
 
1636
- # Remote agent triggers — execute on schedule from anywhere
1637
- /schedule list
1638
- /schedule run <trigger-id>
2036
+ Ralph records loop state so work can be inspected and resumed when supported by the selected provider and local environment.
2037
+
2038
+ ```bash
2039
+ aiwg ralph-status
2040
+ aiwg ralph-attach <loop-id>
2041
+ aiwg ralph-resume <loop-id>
2042
+ aiwg ralph-abort <loop-id>
1639
2043
  ```
1640
2044
 
1641
- ### Ralph Control
2045
+ Use budgets for any loop that may call a remote model or external provider:
1642
2046
 
1643
2047
  ```bash
1644
- /ralph-status # Check current/previous loop status
1645
- /ralph-resume # Resume interrupted loop from last checkpoint
1646
- /ralph-abort # Cancel running loop (optionally revert changes)
1647
- /ralph-memory # View debug memory entries and failure patterns
1648
- /ralph-reflect # View episodic memory and strategy evolution
1649
- /ralph-analytics # Execution metrics and performance history
2048
+ aiwg ralph "Reduce flaky integration tests" \
2049
+ --completion "the flaky-test reproduction passes 10 consecutive runs" \
2050
+ --max-total-tokens 200000 \
2051
+ --max-total-cost 10 \
2052
+ --budget-stop-policy budget-wins
1650
2053
  ```
1651
2054
 
1652
- ### How It Works
2055
+ Long-running automation should still produce reviewable outputs: changed files, reports, evidence, status logs, and
2056
+ the exact checks run. Ralph can continue work within configured limits, but it cannot guarantee a solution, fixed
2057
+ runtime, or provider availability.
1653
2058
 
1654
- Each iteration follows the TAO loop (Thought → Action → Observation):
2059
+ Mission Control builds on the same principle for multiple bounded work items:
1655
2060
 
1656
- ```
1657
- Iteration N:
1658
- 1. THINK — Analyze current state + accumulated learnings from iterations 1..N-1
1659
- 2. ACT — Make changes based on task + debug memory + failure patterns
1660
- 3. VERIFY — Run completion command (tests, build, lint, coverage, etc.)
1661
- 4. LEARN — If verification fails, extract root cause → store in debug memory
1662
- 5. DECIDE — Pass? → Complete. Fail? → Iterate. Max retries? → Escalate to human.
2061
+ ```bash
2062
+ aiwg mc start --name "docs audit follow-up" --max-missions 4
2063
+ aiwg mc dispatch <session-id> \
2064
+ "Validate README command examples" \
2065
+ --completion "all documented commands are current or labeled provider-specific"
2066
+ aiwg mc run <session-id>
2067
+ aiwg mc status <session-id>
2068
+ aiwg mc watch <session-id>
1663
2069
  ```
1664
2070
 
1665
- The debug memory system implements executable feedback: the agent doesn't just retry — it learns *what went wrong* and *why*, then applies that knowledge to the next attempt. After 3 failed attempts at the same root cause, it escalates to a human rather than looping forever.
2071
+ ## RLM — Recursive Context Decomposition
1666
2072
 
1667
- Research foundation: Self-Refine (Madaan et al., NeurIPS 2023), ReAct (Yao et al., ICLR 2023), METR 2025 (recovery capability dominates agentic task success), Reflexion (Shinn et al., 2023).
2073
+ For a bounded batch, specify the parallelism limit explicitly:
1668
2074
 
1669
- ---
2075
+ ```text
2076
+ /rlm-batch "src/components/*.tsx" "Add TypeScript types" --max-parallel 4
2077
+ ```
1670
2078
 
1671
- ## RLM — Recursive Context Decomposition
2079
+ RLM helps with sources that are too large to fit comfortably in one model context. It prepares files into traceable
2080
+ chunks, fans a query out across those chunks, and merges results with links back to the source material.
2081
+
2082
+ ```mermaid
2083
+ flowchart LR
2084
+ A[Source files] --> B[rlm-prep]
2085
+ B --> C[manifest.json]
2086
+ C --> D[rlm-search / fanout]
2087
+ D --> E[ranked findings]
2088
+ E --> F[source-linked answer]
2089
+ C --> G[rlm-cache]
2090
+ ```
1672
2091
 
1673
- Process codebases and documents far beyond any model's context window:
2092
+ Use `rlm-prep` when you want to prepare a file tree once and reuse it for several searches.
1674
2093
 
1675
2094
  ```bash
1676
- # Query: fan-out across files, gather results
1677
- /rlm-query "src/**/*.ts" "Extract all exported interfaces" --model haiku
2095
+ # Prepare source or docs for recursive search
2096
+ aiwg rlm-prep src/ --strategy semantic-boundary --size 200 --overlap 20
2097
+ aiwg rlm-prep docs/ --strategy fixed-count --size 150
1678
2098
 
1679
- # Batch: parallel processing with configurable concurrency
1680
- /rlm-batch "src/components/*.tsx" "Add TypeScript types" --max-parallel 4
2099
+ # Search the prepared source
2100
+ aiwg rlm-search "Where is provider reload status calculated?" \
2101
+ --source .aiwg/rlm-prep/<source-hash>/manifest.json \
2102
+ --depth 3 \
2103
+ --max-parallel 4 \
2104
+ --budget 50000
2105
+
2106
+ # Run a direct fanout query over a manifest or chunks directory
2107
+ aiwg fanout "Summarize every stale quickstart command" \
2108
+ --chunks .aiwg/rlm-prep/<source-hash>/manifest.json \
2109
+ --parallel 4
1681
2110
 
1682
- # Status: monitor decomposition progress
1683
- /rlm-status
2111
+ # Inspect cache state
2112
+ aiwg rlm-status
2113
+ aiwg rlm-cache stats
1684
2114
  ```
1685
2115
 
1686
- The RLM addon decomposes large inputs into chunks, delegates each to a sub-agent, and synthesizes results. Processes 10M+ tokens through recursive delegation.
2116
+ Use `chunk` for a single-file manual workflow:
1687
2117
 
1688
- Research foundation: Recursive Language Models (Zhang, Kraska, Khattab — MIT CSAIL, 2026).
2118
+ ```bash
2119
+ aiwg chunk README.md --size 200 --overlap 20 --format json --output .aiwg/chunks/readme
2120
+ ```
1689
2121
 
1690
- ---
2122
+ RLM is a retrieval and decomposition workflow, not a magic context override. Quality depends on chunk boundaries,
2123
+ source coverage, prompt specificity, model capability, and budget. For high-stakes review, ask for quoted source
2124
+ links, inspect the cited chunks, and rerun targeted searches for disputed claims.
1691
2125
 
1692
2126
  ## Research Foundations
1693
2127
 
1694
- AIWG's architecture is grounded in peer-reviewed research across cognitive science, multi-agent systems, software engineering, and AI safety. Reference summaries live in `docs/references/` (REF-NNN entries), ordered highest to lowest GRADE evidence quality within each category.
2128
+ AIWG draws design ideas from research across cognitive science, multi-agent
2129
+ systems, software engineering, retrieval, provenance, and AI safety. The
2130
+ results cited below belong to the referenced papers or systems; they are not
2131
+ AIWG performance guarantees. Reference summaries live in `docs/references/`
2132
+ (REF-NNN entries). The bibliography groups related design topics.
1695
2133
 
1696
2134
  ### Cognitive Foundations
1697
2135
 
@@ -1705,22 +2143,31 @@ AIWG's architecture is grounded in peer-reviewed research across cognitive scien
1705
2143
  ### Multi-Agent Systems & Orchestration
1706
2144
 
1707
2145
  - Jacobs, R.A. et al. (1991). [Adaptive Mixtures of Local Experts](https://doi.org/10.1162/neco.1991.3.1.79). *Neural Computation*, 3(1), 79–87. (Mixture-of-Experts foundation)
1708
- - Hong, S. et al. (2024). [MetaGPT: Meta Programming for a Multi-Agent Collaborative Framework](https://arxiv.org/abs/2308.00352). *ICLR 2024*. (85.9% HumanEval, SOP-based orchestration)
2146
+ - Hong, S. et al. (2024). [MetaGPT: Meta Programming for a Multi-Agent Collaborative
2147
+ Framework](https://arxiv.org/abs/2308.00352). *ICLR 2024*.
1709
2148
  - Qian, C. et al. (2024). [ChatDev: Communicative Agents for Software Development](https://arxiv.org/abs/2307.07924). *ACL 2024*.
1710
2149
  - Shen, Y. et al. (2023). [HuggingGPT: Solving AI Tasks with ChatGPT and its Friends in HuggingFace](https://arxiv.org/abs/2303.17580). *NeurIPS 2023*.
1711
2150
  - Tao, W. et al. (2024). [MAGIS: LLM-Based Multi-Agent Framework for GitHub Issue Resolution](https://arxiv.org/abs/2403.17927).
1712
- - Zhang, J. et al. (2025). [AFlow: Automating Agentic Workflow Generation](https://arxiv.org/abs/2410.10762). *ICLR 2025 Oral*. (5.7% avg gain over best manual methods)
2151
+ - Zhang, J. et al. (2025). [AFlow: Automating Agentic Workflow Generation](https://arxiv.org/abs/2410.10762). *ICLR
2152
+ 2025 Oral*.
1713
2153
  - Wu, Q. et al. (2023). [AutoGen: Enabling Next-Gen LLM Applications via Multi-Agent Conversation](https://arxiv.org/abs/2308.08155). (Conversational multi-agent framework)
1714
2154
  - Yu, C. et al. (2025). [A Survey on Agent Workflow — Status and Future](https://arxiv.org/abs/2508.01186). (24 systems, 11 metrics)
1715
- - Lodha, D. et al. (2026). [MCP-Diag: A Deterministic, Protocol-Driven Architecture for AI-Native Network Diagnostics](https://arxiv.org/abs/2601.22633). *COMSNETS 2026*. (First production MCP system)
1716
- - Yu, G. (2026). [AdaptOrch: Adaptive Orchestration for Multi-Agent LLM Systems Through Topology-Aware Task Planning](https://arxiv.org/abs/2502.09340). (12–23% improvement across 4 topologies)
2155
+ - Lodha, D. et al. (2026). [MCP-Diag: A Deterministic, Protocol-Driven Architecture for AI-Native Network
2156
+ Diagnostics](https://arxiv.org/abs/2601.22633). *COMSNETS 2026*.
2157
+ - Yu, G. (2026). [AdaptOrch: Adaptive Orchestration for Multi-Agent LLM Systems Through Topology-Aware Task Planning](https://arxiv.org/abs/2502.09340).
1717
2158
  - Gerred (2025). [Multi-Agent Orchestration](https://gerred.github.io/building-an-agentic-system/second-edition/part-iv-advanced-patterns/chapter-10-multi-agent-orchestration.html). Tool isolation, resource boundaries, observable coordination.
1718
2159
  - Falconer, S. (2025). [Event-Driven Multi-Agent Systems](https://www.confluent.io/blog/event-driven-multi-agent-systems/). Confluent. 4 Kafka orchestration patterns.
1719
2160
  - Mario, M. (2025). [Multi-Agent System Patterns: A Unified Guide to Designing Agentic Architectures](https://medium.com/@mjgmario/multi-agent-system-patterns-a-unified-guide-to-designing-agentic-architectures-04bb31ab9c41). 4-dimensional framework.
1720
- - Runkle, S. (2026). [Choosing the Right Multi-Agent Architecture](https://www.blog.langchain.com/choosing-the-right-multi-agent-architecture/). LangChain. Subagents, skills, handoffs, 90.2% improvement stat.
1721
- - Towards Data Science (2025). [Why Your Multi-Agent System Is Failing: Escaping the 17x Error Trap](https://towardsdatascience.com/why-your-multi-agent-system-is-failing-escaping-the-17x-error-trap-of-the-bag-of-agents/). 17.2x error amplification, 4-agent coordination threshold.
2161
+ - Runkle, S. (2026). [Choosing the Right Multi-Agent
2162
+ Architecture](https://www.blog.langchain.com/choosing-the-right-multi-agent-architecture/). LangChain. Subagents,
2163
+ skills, and handoffs.
2164
+ - Towards Data Science (2025). [Why Your Multi-Agent System Is Failing: Escaping the 17x Error
2165
+ Trap](https://towardsdatascience.com/why-your-multi-agent-system-is-failing-escaping-the-17x-error-trap-of-the-bag-of-agents/).
2166
+ Coordination failure analysis.
1722
2167
  - NexAI Tech (2025). [Multi-AI Agent Architecture Patterns for Scale](https://nexaitech.com/multi-ai-agent-architecutre-patterns-for-scale/). Enterprise 5-layer architecture, 3 orchestration patterns.
1723
- - Wexford, E. (2026). [How to Build Multi-Agent Systems: Complete 2026 Guide](https://dev.to/eira-wexford/how-to-build-multi-agent-systems-complete-2026-guide-1io6). DEV Community. 3–7 agents optimal sizing.
2168
+ - Wexford, E. (2026). [How to Build Multi-Agent Systems: Complete 2026
2169
+ Guide](https://dev.to/eira-wexford/how-to-build-multi-agent-systems-complete-2026-guide-1io6). DEV Community.
2170
+ Multi-agent design guidance.
1724
2171
 
1725
2172
  ### Reasoning & Planning
1726
2173
 
@@ -1730,11 +2177,13 @@ AIWG's architecture is grounded in peer-reviewed research across cognitive scien
1730
2177
  - Yao, S. et al. (2023). [Tree of Thoughts: Deliberate Problem Solving with Large Language Models](https://arxiv.org/abs/2305.10601). *NeurIPS 2023*.
1731
2178
  - Zhou, A. et al. (2024). [Language Agent Tree Search Unifies Reasoning, Acting, and Planning in Language Models](https://arxiv.org/abs/2310.04406). *ICML 2024*.
1732
2179
  - Kojima, T. et al. (2022). [Large Language Models are Zero-Shot Reasoners](https://arxiv.org/abs/2205.11916). *NeurIPS 2022*. ("Let's think step by step")
1733
- - Liu, Z. et al. (2026). [Exploratory Memory-Augmented LLM Agent via Hybrid On- and Off-Policy Optimization (EMPO²)](https://arxiv.org/abs/2602.23008). *ICLR 2026*. (128.6% over GRPO on ScienceWorld)
2180
+ - Liu, Z. et al. (2026). [Exploratory Memory-Augmented LLM Agent via Hybrid On- and Off-Policy Optimization
2181
+ (EMPO²)](https://arxiv.org/abs/2602.23008). *ICLR 2026*.
1734
2182
 
1735
2183
  ### Self-Correction & Iterative Refinement
1736
2184
 
1737
- - Madaan, A. et al. (2023). [Self-Refine: Iterative Refinement with Self-Feedback](https://arxiv.org/abs/2303.17651). *NeurIPS 2023*. (+4.2% HumanEval, −63% revision cost)
2185
+ - Madaan, A. et al. (2023). [Self-Refine: Iterative Refinement with Self-Feedback](https://arxiv.org/abs/2303.17651).
2186
+ *NeurIPS 2023*.
1738
2187
  - Shinn, N. et al. (2023). [Reflexion: Language Agents with Verbal Reinforcement Learning](https://arxiv.org/abs/2303.11366). *NeurIPS 2023*.
1739
2188
 
1740
2189
  ### Stage-Gate, SDLC & Traceability
@@ -1746,45 +2195,49 @@ AIWG's architecture is grounded in peer-reviewed research across cognitive scien
1746
2195
  ### Software Engineering & Agent-Computer Interface
1747
2196
 
1748
2197
  - Jimenez, C.E. et al. (2024). [SWE-bench: Can Language Models Resolve Real-world GitHub Issues?](https://www.swebench.com). *ICLR 2024*.
1749
- - Wang, X. et al. (2024). [Executable Code Actions Elicit Better LLM Agents (CodeAct)](https://arxiv.org/abs/2402.01030). *ICML 2024*. (Up to 20% higher success rate)
1750
- - Yang, J. et al. (2024). [SWE-agent: Agent-Computer Interfaces Enable Automated Software Engineering](https://arxiv.org/abs/2405.15793). *NeurIPS 2024*. (12.47% SWE-bench)
1751
- - Laurent, A. (2025). [A Comparison of AI Code Assistants for Large Codebases](https://intuitionlabs.ai/articles/ai-code-assistants-large-codebases). IntuitionLabs. (62% AI code contains flaws)
1752
- - Augment Code (2025). [AI Coding Assistants for Large Codebases: A Complete Guide](https://www.augmentcode.com/tools/ai-coding-assistants-for-large-codebases-a-complete-guide). (73% compile locally but violate patterns)
2198
+ - Wang, X. et al. (2024). [Executable Code Actions Elicit Better LLM Agents
2199
+ (CodeAct)](https://arxiv.org/abs/2402.01030). *ICML 2024*.
2200
+ - Yang, J. et al. (2024). [SWE-agent: Agent-Computer Interfaces Enable Automated Software
2201
+ Engineering](https://arxiv.org/abs/2405.15793). *NeurIPS 2024*.
2202
+ - Laurent, A. (2025). [A Comparison of AI Code Assistants for Large
2203
+ Codebases](https://intuitionlabs.ai/articles/ai-code-assistants-large-codebases). IntuitionLabs.
2204
+ - Augment Code (2025). [AI Coding Assistants for Large Codebases: A Complete Guide](https://www.augmentcode.com/tools/ai-coding-assistants-for-large-codebases-a-complete-guide).
1753
2205
  - AlgoMaster (2025). [How to Use AI Effectively in Large Codebases](https://blog.algomaster.io/p/using-ai-effectively-in-large-codebases). Retrieval as bottleneck framing.
1754
2206
 
1755
2207
  ### Context Engineering & Memory
1756
2208
 
1757
2209
  - Liu, N.F. et al. (2024). [Lost in the Middle: How Language Models Use Long Contexts](https://arxiv.org/abs/2307.03172). *TACL* 12, 157–173. doi:10.1162/tacl_a_00638
1758
2210
  - Dai, Y. et al. (2025). [Pretraining Context Compressor for Large Language Models with Embedding-Based Memory](https://aclanthology.org/2025.acl-long.1394.pdf). *ACL 2025*.
1759
- - Kang, M. et al. (2025). [ACON: Optimizing Context Compression for Long-Horizon LLM Agents](https://arxiv.org/abs/2510.00615). (26–54% peak token reduction, >95% accuracy preserved)
1760
- - Liu, F. & Qiu, H. (2025). [Context Cascade Compression (C3): Exploring the Upper Limits of Text Compression](https://arxiv.org/abs/2511.15244). (98% precision at 20x compression)
2211
+ - Kang, M. et al. (2025). [ACON: Optimizing Context Compression for Long-Horizon LLM Agents](https://arxiv.org/abs/2510.00615).
2212
+ - Liu, F. & Qiu, H. (2025). [Context Cascade Compression (C3): Exploring the Upper Limits of Text Compression](https://arxiv.org/abs/2511.15244).
1761
2213
  - Vasilopoulos, A. (2026). [Codified Context: Infrastructure for AI Agents in a Complex Codebase](https://arxiv.org/abs/2602.20478). (Three-tier context infrastructure: constitution + 19 agents + 34-doc KB)
1762
2214
  - Ostby, D.L. (2025). [Stingy Context: Compressing Code Context for Cost-Effective AI Development Assistance](https://arxiv.org/abs/2512.15504). (TREEFRAG, 18:1 compression ratio)
1763
2215
  - Anthropic Applied AI Team (2026). [Effective Context Engineering for AI Agents](https://www.anthropic.com/engineering/effective-context-engineering-for-ai-agents). Anthropic Engineering Blog.
1764
- - Huang, J.Y. et al. (2026). [Do LLMs Benefit From Their Own Words?](https://arxiv.org/abs/2602.24287) (36.4% of multi-turn prompts self-contained; up to 10x context reduction)
2216
+ - Huang, J.Y. et al. (2026). [Do LLMs Benefit From Their Own Words?](https://arxiv.org/abs/2602.24287)
1765
2217
  - Böckeler, B. (2026). [Context Engineering for Coding Agents](https://martinfowler.com/articles/exploring-gen-ai/context-engineering-coding-agents.html). Martin Fowler's Blog. Two-category framework.
1766
- - Haseeb, M. (2025). [Context Engineering for Multi-Agent LLM Code Assistants](https://arxiv.org/abs/2508.08322). (80% vs 40% single-shot success)
1767
- - Verma, N. (2026). [Focus Agent: LLM Agent with Active Context Compression for SWE-Bench](https://arxiv.org/abs/2501.09067). (22.7% token reduction via consolidate/withdraw)
2218
+ - Haseeb, M. (2025). [Context Engineering for Multi-Agent LLM Code Assistants](https://arxiv.org/abs/2508.08322).
2219
+ - Verma, N. (2026). [Focus Agent: LLM Agent with Active Context Compression for SWE-Bench](https://arxiv.org/abs/2501.09067).
1768
2220
  - Zylos Research (2026). [Long-Running AI Agents and Task Decomposition](https://zylos.ai/research/2026-01-16-long-running-ai-agents). (35-min degradation threshold, Planner-Worker model)
1769
- - Zylos Research (2026). [LLM Context Window Management and Long-Context Strategies](https://zylos.ai/research/2026-01-19-llm-context-management). (Lost-in-Middle persists, TTT-E2E 35× speedup)
2221
+ - Zylos Research (2026). [LLM Context Window Management and Long-Context Strategies](https://zylos.ai/research/2026-01-19-llm-context-management).
1770
2222
 
1771
2223
  ### Agent Memory & Knowledge Systems
1772
2224
 
1773
2225
  - Laird, J.E. et al. (1987). [SOAR: An Architecture for General Intelligence](https://doi.org/10.1016/0004-3702(87)90050-6). *Artificial Intelligence*, 33(1), 1–64.
1774
2226
  - Anderson, J.R. et al. (2004). [An Integrated Theory of the Mind (ACT-R)](https://doi.org/10.1037/0033-295X.111.4.1036). *Psychological Review*, 111(4), 1036–1060.
1775
2227
  - Park, J.S. et al. (2023). [Generative Agents: Interactive Simulacra of Human Behavior](https://arxiv.org/abs/2304.03442). *UIST 2023*. doi:10.1145/3586183.3606763
1776
- - Xu, W. et al. (2025). [A-MEM: Agentic Memory for LLM Agents](https://arxiv.org/abs/2502.12110). (Zettelkasten-inspired, 85–93% token reduction)
2228
+ - Xu, W. et al. (2025). [A-MEM: Agentic Memory for LLM Agents](https://arxiv.org/abs/2502.12110).
1777
2229
  - Hu, Y. et al. (2025). [Memory in the Age of AI Agents: A Survey](https://arxiv.org/abs/2512.13564). (Surveys 100+ implementations, forms-functions-dynamics framework)
1778
- - Rezazadeh, A. et al. (2025). [Collaborative Memory: Multi-User Memory Sharing in LLM Agents with Dynamic Access Control](https://arxiv.org/abs/2505.18279). (61% resource reduction)
2230
+ - Rezazadeh, A. et al. (2025). [Collaborative Memory: Multi-User Memory Sharing in LLM Agents with Dynamic Access Control](https://arxiv.org/abs/2505.18279).
1779
2231
  - Yuen, S. et al. (2025). [Intrinsic Memory Agents: Heterogeneous Multi-Agent LLM Systems through Structured Contextual Memory](https://arxiv.org/abs/2508.08997). Role-aligned heterogeneous memory.
1780
2232
  - Graves, A., Wayne, G. & Danihelka, I. (2014). [Neural Turing Machines](https://arxiv.org/abs/1410.5401). External memory architectures.
1781
2233
  - Packer, C. et al. (2023). [MemGPT: Towards LLMs as Operating Systems](https://arxiv.org/abs/2310.08560). OS-inspired virtual context paging.
1782
2234
  - Yu, Z. et al. (2026). [Multi-Agent Memory from a Computer Architecture Perspective](https://arxiv.org/abs/2603.10062). *Architecture 2.0 '26*. Three-layer I/O-cache-memory hierarchy.
1783
- - Chhikara, P. et al. (2025). [Mem0: Building Production-Ready AI Agents with Scalable Long-Term Memory](https://arxiv.org/abs/2504.19413). (26% accuracy gain, 91% latency reduction)
2235
+ - Chhikara, P. et al. (2025). [Mem0: Building Production-Ready AI Agents with Scalable Long-Term Memory](https://arxiv.org/abs/2504.19413).
1784
2236
 
1785
2237
  ### Recursive Context Decomposition
1786
2238
 
1787
- - Zhang, A.L., Kraska, T. & Khattab, O. (2026). [Recursive Language Models](https://arxiv.org/abs/2512.24601). *arXiv:2512.24601*. MIT CSAIL. (10M+ token processing, up to 3x cheaper than summarization)
2239
+ - Zhang, A.L., Kraska, T. & Khattab, O. (2026). [Recursive Language Models](https://arxiv.org/abs/2512.24601).
2240
+ *arXiv:2512.24601*. MIT CSAIL.
1788
2241
 
1789
2242
  ### Provenance, Reproducibility & Research Management
1790
2243
 
@@ -1792,8 +2245,8 @@ AIWG's architecture is grounded in peer-reviewed research across cognitive scien
1792
2245
  - W3C (2013). [PROV-DM: The PROV Data Model](https://www.w3.org/TR/prov-dm/). W3C Recommendation.
1793
2246
  - CCSDS (2024). [Reference Model for an Open Archival Information System (OAIS)](https://public.ccsds.org/Pubs/650x0m2.pdf). ISO 14721. (Digital preservation lifecycle)
1794
2247
  - GRADE Working Group (2004–present). [GRADE Handbook](https://www.gradeworkinggroup.org/). Evidence quality assessment. Adopted by WHO, Cochrane, NICE, and 100+ organizations.
1795
- - Schmidgall, S. et al. (2025). [Agent Laboratory: Using LLM Agents as Research Assistants](https://arxiv.org/abs/2501.04227). (84% cost reduction)
1796
- - Sureshkumar, V. et al. (2026). [R-LAM: Towards Reproducibility in Large Action Model Workflows](https://arxiv.org/abs/2601.09749). (47% of workflows non-reproducible without constraints)
2248
+ - Schmidgall, S. et al. (2025). [Agent Laboratory: Using LLM Agents as Research Assistants](https://arxiv.org/abs/2501.04227).
2249
+ - Sureshkumar, V. et al. (2026). [R-LAM: Towards Reproducibility in Large Action Model Workflows](https://arxiv.org/abs/2601.09749).
1797
2250
  - ServiceNow Research (2025). LitLLM for Scientific Literature Reviews. RAG-based literature review, no hallucination approach.
1798
2251
 
1799
2252
  ### AI Safety & Failure Modes
@@ -1833,81 +2286,111 @@ AIWG's architecture is grounded in peer-reviewed research across cognitive scien
1833
2286
 
1834
2287
  ### Constrained Generation & Output Validation
1835
2288
 
1836
- - Beurer-Kellner, L., Fischer, M. & Vechev, M. (2023). [Prompting Is Programming: A Query Language for Large Language Models (LMQL)](https://arxiv.org/abs/2212.06094). *PLDI 2023*. doi:10.1145/3591300 (26–85% token reduction)
1837
- - Willard, B.T. & Louf, R. (2023). [Efficient Guided Generation for Large Language Models (Outlines)](https://arxiv.org/abs/2307.09702). (0% parse failures by construction)
1838
- - Lhoest, Q. & Turuta, M. (2024). [Structured Generation with Outlines](https://huggingface.co/blog/outlines-structured-generation). Hugging Face Blog. (1.5–3x speedup)
2289
+ - Beurer-Kellner, L., Fischer, M. & Vechev, M. (2023). [Prompting Is Programming: A Query Language for Large Language
2290
+ Models (LMQL)](https://arxiv.org/abs/2212.06094). *PLDI 2023*. doi:10.1145/3591300
2291
+ - Willard, B.T. & Louf, R. (2023). [Efficient Guided Generation for Large Language Models (Outlines)](https://arxiv.org/abs/2307.09702).
2292
+ - Lhoest, Q. & Turuta, M. (2024). [Structured Generation with
2293
+ Outlines](https://huggingface.co/blog/outlines-structured-generation). Hugging Face Blog.
1839
2294
  - Gerganov, G. et al. (2024). [Grammar-Based Sampling (GBNF) — llama.cpp](https://github.com/ggerganov/llama.cpp/blob/master/grammars/README.md). Context-free grammar constrained sampling.
1840
2295
 
1841
2296
  ### LLM Serving & Local Deployment
1842
2297
 
1843
- - Yu, G. et al. (2022). [Orca: A Distributed Serving System for Transformer-Based Generative Models](https://www.usenix.org/conference/osdi22/presentation/yu). *OSDI '22*. (36.9x throughput improvement)
1844
- - Kwon, W. et al. (2023). [Efficient Memory Management for Large Language Model Serving with PagedAttention](https://arxiv.org/abs/2309.06180). *SOSP '23*. UC Berkeley. (2–4x throughput vs HuggingFace)
2298
+ - Yu, G. et al. (2022). [Orca: A Distributed Serving System for Transformer-Based Generative
2299
+ Models](https://www.usenix.org/conference/osdi22/presentation/yu). *OSDI '22*.
2300
+ - Kwon, W. et al. (2023). [Efficient Memory Management for Large Language Model Serving with
2301
+ PagedAttention](https://arxiv.org/abs/2309.06180). *SOSP '23*. UC Berkeley.
1845
2302
  - Ollama Team (2024). [Ollama Concurrent Requests and Performance FAQ](https://github.com/ollama/ollama/blob/main/docs/faq.md). `OLLAMA_NUM_PARALLEL` configuration guidance.
1846
2303
 
1847
2304
  ### MCP & Agentic Standards
1848
2305
 
1849
2306
  - Agentic AI Foundation / Linux Foundation (2025). [Model Context Protocol Specification 2025-11-25](https://modelcontextprotocol.io/specification/2025-11-25). (Tool integration protocol)
1850
2307
 
1851
- Full research background, citations, and methodology: [docs/research/](docs/research/)
1852
-
1853
2308
  ---
1854
2309
 
1855
2310
  ## Why AIWG
1856
2311
 
2312
+ AIWG is for people who already use AI assistants and want the work to survive
2313
+ past one conversation. It gives the assistant project-readable instructions,
2314
+ specialist roles, repeatable workflows, and a place to save plans, findings,
2315
+ and decisions.
2316
+
1857
2317
  ### For Individual Developers
1858
2318
 
1859
- **Turn your AI coding assistant from a stateless autocomplete into a project-aware development partner.** Without AIWG, every time your AI assistant restarts you lose all context. With AIWG, the `.aiwg/` directory maintains requirements, architecture decisions, test strategies, and project history across sessions. The agent loop means you can hand off complex multi-step tasks ("migrate this module to TypeScript") and walk away — the agent iterates until completion or escalates when stuck.
2319
+ Use AIWG when a coding task needs project context, not just a one-off answer.
2320
+ The `.aiwg/` directory can hold requirements, architecture notes, test plans,
2321
+ reviews, and task reports. Later sessions can read those artifacts before
2322
+ changing code. Agent-loop workflows can continue a bounded task until the
2323
+ named verification check passes, a limit is reached, or human input is needed.
1860
2324
 
1861
2325
  ### For Engineering Teams
1862
2326
 
1863
- **Standardize how your team works with AI across 14 provider integrations.** Whether your team uses Antigravity, Claude Code, Codex, Copilot, Cursor, Pi, or Warp, everyone can receive the same source workflows adapted to the provider's supported surfaces. The 35 enforcement rules prevent common AI mistakes: deleting tests to make them pass, fabricating citations, hard-coding tokens, or silently dropping features. Human-in-the-loop gates at phase transitions ensure no AI output reaches production without human review.
2327
+ Use AIWG to keep shared instructions and review artifacts with the repository.
2328
+ Teams can deploy the same source workflows to the provider surfaces they use,
2329
+ then verify each provider handoff separately. Human gates and review steps help
2330
+ make important changes visible before they move forward; they do not replace
2331
+ human code review, testing, security review, or release approval.
1864
2332
 
1865
2333
  ### For Platform Engineers
1866
2334
 
1867
- **Deploy consistent AI-augmented workflows across your organization.** AIWG's extension system lets you create custom agents, commands, and skills specific to your domain, then deploy them to any supported platform. The scaffolding commands (`aiwg add-agent`, `aiwg scaffold-addon`, `aiwg scaffold-framework`) make it easy to build and distribute organizational capabilities.
2335
+ Use AIWG when you need reusable AI workflows across projects or teams. The
2336
+ extension system supports project-local rules, skills, agents, addons, and
2337
+ frameworks. Scaffolding commands help create those components, and discovery
2338
+ helps agents find the resulting capabilities without requiring users to
2339
+ memorize the catalog.
1868
2340
 
1869
2341
  ### For Researchers
1870
2342
 
1871
- **Standards-aligned implementation of multi-agent systems with full provenance.** AIWG operationalizes FAIR Principles (G20, EU, NIH endorsement), W3C PROV for provenance tracking, and GRADE methodology for evidence quality. The research framework manages paper discovery, citation integrity, and archival lifecycles. All 138+ research paper citations are grounded in verified sources — no hallucinated references.
2343
+ Use AIWG to organize source notes, citations, provenance records, and research
2344
+ artifacts. The research framework uses ideas from FAIR, W3C PROV, GRADE, and
2345
+ OAIS to make research work easier to inspect and maintain. Those alignments are
2346
+ documentation and workflow structures, not a certification that every generated
2347
+ note is correct.
1872
2348
 
1873
2349
  ### For Security Teams
1874
2350
 
1875
- **Forensics-grade investigation workflows and security review automation.** The Forensics Complete framework provides 13 specialized agents for digital forensics and incident response, with NIST SP 800-86 evidence handling, MITRE ATT&CK mapping, Sigma rule hunting, and STIX 2.1 IOC formatting. Security review cycles integrate into the SDLC with automated threat modeling and vulnerability management.
2351
+ Use AIWG for structured security reviews, incident investigation notes, and
2352
+ forensics-oriented workflows. The forensics framework includes target
2353
+ profiling, triage, acquisition, timeline, IOC, and reporting workflows with
2354
+ references to NIST SP 800-86, MITRE ATT&CK, Sigma, and STIX. Investigations
2355
+ still require authorized access, evidence handling discipline, and human
2356
+ review.
1876
2357
 
1877
2358
  ---
1878
2359
 
1879
2360
  ## AIWG vs Manual AI Workflows
1880
2361
 
1881
- | Capability | Without AIWG | With AIWG |
1882
- |-----------|-------------|-----------|
1883
- | **Context persistence** | Lost on every restart | `.aiwg/` survives across sessions |
1884
- | **Multi-agent coordination** | Manual prompt switching | Orchestrated parallel reviews with synthesis |
1885
- | **Quality enforcement** | Hope for the best | 35 rules auto-enforced (anti-laziness, token security, citation integrity) |
1886
- | **Error recovery** | Start over | Agent loop iterates with learned debug memory |
1887
- | **Long-running tasks** | Babysit the terminal | External agent loop runs 6-8+ hours crash-resilient |
1888
- | **Traceability** | Grep and hope | @-mention system with bidirectional linking |
1889
- | **Reproducibility** | Non-deterministic | Strict mode (temperature=0), checkpoints, validation |
1890
- | **Platform switching** | Rewrite all prompts | `--provider copilot` deploys identical workflows |
1891
- | **Citation integrity** | AI may hallucinate | Retrieval-first architecture, GRADE-assessed sources only |
1892
- | **Phase management** | Ad-hoc | Stage-gate with human approval at transitions |
2362
+ | Task | Manual AI workflow | AIWG mechanism | Limit to keep in view |
2363
+ |------|--------------------|----------------|-----------------------|
2364
+ | Carry context into another session | Paste summaries, links, and decisions again | Save project artifacts under `.aiwg/` and route later work through them | The assistant still has to read and interpret the right artifacts |
2365
+ | Coordinate review perspectives | Ask separate prompts and merge notes by hand | Use specialist roles and review workflows that produce a combined artifact | Reviewers can share wrong assumptions if the input context is wrong |
2366
+ | Keep task scope visible | Track instructions in chat history | Use workflows, rules, and saved task outputs with explicit scope | Prompt-level rules are not hard technical enforcement by themselves |
2367
+ | Recover from a failed attempt | Restart from the last visible message | Use loop state, checkpoints, reports, and bounded retries where configured | Recovery can still stop on missing tools, unclear goals, or failing tests |
2368
+ | Work across AI tools | Rewrite prompts for each provider | Deploy provider-specific files from the same source workflows | Provider capabilities, permissions, and reload behavior differ |
2369
+ | Trace decisions to work products | Search notes manually | Use artifacts, mentions, indexes, and provenance records | Links can drift and need validation |
2370
+ | Review citations and research notes | Trust generated citations or manually inspect each one | Store source records, notes, citations, and quality assessments together | Source-grounding reduces risk; it does not prove every claim is correct |
2371
+ | Move through project phases | Keep phase criteria in a checklist | Use stage-gate workflows with human approval points | Gates reflect configured criteria and available evidence |
1893
2372
 
1894
2373
  ---
1895
2374
 
1896
- ## Standards & Compliance
1897
-
1898
- | Standard | How AIWG Uses It |
1899
- |----------|-----------------|
1900
- | **FAIR Principles** (G20, EU, NIH) | Findable, Accessible, Interoperable, Reusable artifact management |
1901
- | **W3C PROV** | Provenance tracking for all generated artifacts |
1902
- | **GRADE** (WHO, Cochrane, NICE) | Evidence quality assessment for research citations |
1903
- | **OAIS** (ISO 14721) | Archival lifecycle management for research corpus |
1904
- | **NIST SP 800-86** | Digital forensics evidence handling |
1905
- | **MITRE ATT&CK** | Threat technique mapping in forensics framework |
1906
- | **STIX 2.1** | Indicator of Compromise formatting |
1907
- | **Sigma Rules** | Threat detection rule format |
1908
- | **IEEE 830** | Requirements specification traceability |
1909
- | **MCP** (Linux Foundation) | Model Context Protocol for tool integration |
1910
- | **CalVer** | Calendar versioning (YYYY.M.PATCH) |
2375
+ ## Standards Alignment
2376
+
2377
+ AIWG uses standards and established methods as design references. This section
2378
+ describes the intended mapping; it is not a compliance guarantee, audit
2379
+ attestation, or substitute for domain-specific review.
2380
+
2381
+ | Standard or method | How AIWG uses it |
2382
+ |--------------------|------------------|
2383
+ | **FAIR Principles** | Artifact and research-corpus structure that favors findable, accessible, interoperable, and reusable records |
2384
+ | **W3C PROV** | Provenance records for selected generated artifacts and derived outputs |
2385
+ | **GRADE** | Evidence-quality language and review patterns for research citations |
2386
+ | **OAIS** (ISO 14721) | Archival lifecycle concepts for research and media corpus handling |
2387
+ | **NIST SP 800-86** | Digital-forensics evidence-handling references in forensics workflows |
2388
+ | **MITRE ATT&CK** | Threat-technique mapping references for security and forensics analysis |
2389
+ | **STIX 2.1** | Indicator-of-compromise formatting references |
2390
+ | **Sigma Rules** | Threat-detection rule format references |
2391
+ | **IEEE 830** | Requirements-specification and traceability influence for SDLC artifacts |
2392
+ | **MCP** | Model Context Protocol integration for tool-based AI workflows |
2393
+ | **CalVer** | Calendar versioning format for AIWG releases |
1911
2394
 
1912
2395
  ---
1913
2396
 
@@ -1915,25 +2398,27 @@ Full research background, citations, and methodology: [docs/research/](docs/rese
1915
2398
 
1916
2399
  ### Getting Started
1917
2400
 
1918
- - **[Quick Start Guide](docs/quickstart.md)** — Install and deploy in minutes
1919
- - **[Prerequisites](docs/getting-started/prerequisites.md)** — Node.js, AI platforms, OS support
2401
+ - **[Quick Start Guide](docs/quickstart.md)** — Connect a project and get a saved first result
2402
+ - **[Install, Connect, and Verify](docs/getting-started/install-connect-verify.md)** — Canonical first-time setup and
2403
+ repair path
2404
+ - **[Prerequisites](docs/getting-started/prerequisites.md)** — Node.js, AI platforms, and operating-system notes
1920
2405
  - **[Agent and Operator Reference](docs/agents/README.md)** — deterministic
1921
2406
  commands, flags, outputs, and recovery contracts for agents and advanced
1922
2407
  operators
1923
2408
 
1924
2409
  ### Customize
1925
2410
 
1926
- - **[Make AIWG Yours](docs/customization/README.md)** — Personal rules, agents, and skills that go live immediately
1927
- - **[Customization Examples](docs/customization/examples.md)** — 5 concrete examples of what people actually customize
2411
+ - **[Make AIWG Yours](docs/customization/README.md)** — Project-local rules, agents, and skills
2412
+ - **[Customization Examples](docs/customization/examples.md)** — Concrete examples of what teams customize
1928
2413
  - **[Fork Workflow](docs/customization/fork-workflow.md)** — Upstream sync, contributing back, the ownership model
1929
2414
 
1930
2415
  ### By Audience
1931
2416
 
1932
2417
  **Practitioners:**
1933
2418
 
1934
- - [Quick Start Guide](docs/quickstart.md) — Hands-on workflows
1935
- - [Agent Loop Guide](docs/ralph-guide.md) — Iterative execution with crash recovery
1936
- - [Platform Guides](docs/integrations/) — 5-10 minute setup per platform
2419
+ - [Quick Start Guide](docs/quickstart.md) — Hands-on first workflow
2420
+ - [Agent Loop Guide](docs/ralph-guide.md) — Iterative execution with explicit completion checks
2421
+ - [Platform Guides](docs/integrations/) — Provider-specific setup and handoff details
1937
2422
 
1938
2423
  **Technical Leaders:**
1939
2424
 
@@ -1945,27 +2430,36 @@ Full research background, citations, and methodology: [docs/research/](docs/rese
1945
2430
 
1946
2431
  - [Research Background](docs/research/) — Literature review and citations
1947
2432
  - [Glossary](docs/research/glossary.md) — Professional terminology mapping
1948
- - [Production-Grade Guide](docs/production-grade-guide.md) — Failure mode mitigation
2433
+ - [Production-Grade Guide](docs/frameworks/sdlc-complete/production-grade-guide.md) — Failure mode mitigation patterns
1949
2434
 
1950
2435
  ### Platform Guides
1951
2436
 
1952
- - **[Claude Code](docs/integrations/claude-code-quickstart.md)** — 5-10 min setup
1953
- - **[Warp Terminal](docs/integrations/warp-terminal-quickstart.md)** — 3-5 min setup
1954
- - **[Factory AI](docs/integrations/factory-quickstart.md)** — 5-10 min setup
1955
- - **[Cursor](docs/integrations/cursor-quickstart.md)** — 5-10 min setup
1956
- - **[All Integrations](docs/integrations/)**
2437
+ - **[Claude Code](docs/integrations/claude-code-quickstart.md)** — Claude Code setup and handoff
2438
+ - **[OpenAI Codex](docs/integrations/codex-quickstart.md)** — Codex setup and handoff
2439
+ - **[GitHub Copilot](docs/integrations/copilot-quickstart.md)** — Copilot setup and handoff
2440
+ - **[Warp Terminal](docs/integrations/warp-terminal-quickstart.md)** — Warp setup and handoff
2441
+ - **[Factory AI](docs/integrations/factory-quickstart.md)** — Factory setup and handoff
2442
+ - **[Cursor](docs/integrations/cursor-quickstart.md)** — Cursor setup and handoff
2443
+ - **[All Integrations](docs/integrations/)** — Provider guide directory
1957
2444
 
1958
2445
  ### Framework Documentation
1959
2446
 
1960
- - **[SDLC Framework](agentic/code/frameworks/sdlc-complete/README.md)** — 98 agents, phase workflows, quality gates
2447
+ - **[SDLC Framework](agentic/code/frameworks/sdlc-complete/README.md)** — Phase workflows, quality gates, and
2448
+ development artifacts
1961
2449
  - **[Forensics Complete](agentic/code/frameworks/forensics-complete/README.md)** — DFIR investigation workflows
1962
- - **[Marketing Kit](agentic/code/frameworks/media-marketing-kit/README.md)** — 37 agents, campaign lifecycle
2450
+ - **[Marketing Kit](agentic/code/frameworks/media-marketing-kit/README.md)** — Campaign lifecycle, content, brand, and
2451
+ review workflows
1963
2452
  - **[Media Curator](agentic/code/frameworks/media-curator/README.md)** — Media archive management
1964
- - **[Research Complete](agentic/code/frameworks/research-complete/README.md)** — 8-stage research pipeline
2453
+ - **[Research Complete](agentic/code/frameworks/research-complete/README.md)** — Research pipeline, source notes,
2454
+ citation, and archive workflows
2455
+ - **[Knowledge Base](agentic/code/frameworks/knowledge-base/README.md)** — Source ingest, wiki pages, and corpus health
2456
+ - **[Ops Complete](agentic/code/frameworks/ops-complete/README.md)** — Runbooks, infrastructure reviews, and
2457
+ operational workflows
1965
2458
 
1966
2459
  ### Extension System
1967
2460
 
1968
- AIWG's unified extension system enables dynamic discovery, semantic search, and cross-platform deployment:
2461
+ AIWG's extension system supports discovery, semantic search, and
2462
+ cross-platform deployment for project-local and packaged capabilities:
1969
2463
 
1970
2464
  - **[Extension System Overview](docs/extensions/overview.md)** — Architecture and capabilities
1971
2465
  - **[Creating Extensions](docs/extensions/creating-extensions.md)** — Build custom agents, commands, skills
@@ -1974,26 +2468,28 @@ AIWG's unified extension system enables dynamic discovery, semantic search, and
1974
2468
  ### Advanced Topics
1975
2469
 
1976
2470
  - **[Agent Loop](docs/ralph-guide.md)** — Iterative task execution with crash recovery
1977
- - **[RLM Addon](agentic/code/addons/rlm/README.md)** — Recursive context decomposition for 10M+ token processing
1978
- - **[Daemon Mode](docs/daemon-guide.md)** — Background file watching, cron scheduling, IPC
1979
- - **[Messaging Integration](docs/messaging-guide.md)** — Bidirectional Slack, Discord, and Telegram bots
1980
- - **[MCP Server](docs/mcp/)** — Model Context Protocol integration
1981
- - **[Agent Design Bible](docs/AGENT-DESIGN.md)** — 10 Golden Rules for agent creation
2471
+ - **[RLM Addon](agentic/code/addons/rlm/README.md)** — Recursive context decomposition
2472
+ - **[External Automation](docs/getting-started/daemon-and-automation.md)** — Current automation boundaries and
2473
+ external-job contracts
2474
+ - **[MCP Server](docs/mcp/README.md)** — Model Context Protocol integration
2475
+ - **[Agent Design](docs/frameworks/sdlc-complete/agent-design.md)** — Agent creation guidance
1982
2476
  - **[YAML Metalanguage](agentic/code/frameworks/sdlc-complete/schemas/metalanguage/)** — Declarative workflow schemas
2477
+ - **[Usage Notes](docs/usage-notes.md)** — Rate-limit and usage guidance
1983
2478
 
1984
2479
  ---
1985
2480
 
1986
2481
  ## Contributing
1987
2482
 
1988
- We welcome contributions! See [CONTRIBUTING.md](CONTRIBUTING.md) for guidelines.
2483
+ Contributions are welcome. See [CONTRIBUTING.md](CONTRIBUTING.md) for the
2484
+ project guidelines.
1989
2485
 
1990
2486
  **Quick contributions:**
1991
2487
 
1992
- - Found an AI pattern? [Open an issue](https://github.com/jmagly/aiwg/issues/new)
1993
- - Have a better rewrite? Submit a PR to `examples/`
1994
- - Want to add an agent? Use `aiwg add-agent` or see `docs/development/agent-template.md`
1995
- - Want to add a skill? Use `aiwg add-skill`
1996
- - Want to create an addon? Use `aiwg scaffold-addon`
2488
+ - Found a bug or confusing workflow? [Open an issue](https://github.com/jmagly/aiwg/issues/new).
2489
+ - Have a documentation improvement? Submit a PR with the source file and the behavior it clarifies.
2490
+ - Want to add an agent? Use `aiwg add-agent` or see `docs/development/agent-template.md`.
2491
+ - Want to add a skill? Use `aiwg add-skill`.
2492
+ - Want to create an addon? Use `aiwg scaffold-addon`.
1997
2493
 
1998
2494
  ---
1999
2495
 
@@ -2004,13 +2500,16 @@ We welcome contributions! See [CONTRIBUTING.md](CONTRIBUTING.md) for guidelines.
2004
2500
  - **Telegram:** [Join Group](https://t.me/+oJg9w2lE6A5lOGFh)
2005
2501
  - **Issues:** [GitHub Issues](https://github.com/jmagly/aiwg/issues)
2006
2502
  - **Discussions:** [GitHub Discussions](https://github.com/jmagly/aiwg/discussions)
2007
- - **Security:** Report vulnerabilities per [SECURITY.md](SECURITY.md) — do not file public issues.
2503
+ - **Security:** Report vulnerabilities per [SECURITY.md](SECURITY.md); do not file public issues for private
2504
+ vulnerability reports.
2008
2505
 
2009
2506
  ---
2010
2507
 
2011
2508
  ## Badges
2012
2509
 
2013
- Building on AIWG? Show it. Grab a **Built With AIWG** / **Powered By AIWG** badge — light and dark, hosted and free to hot-link (no need to copy the image into your repo):
2510
+ Building on AIWG? You can use a **Built With AIWG** or **Powered By AIWG**
2511
+ badge. Hosted image links are available, and projects can also copy the image
2512
+ if they prefer to avoid hot-linking:
2014
2513
 
2015
2514
  [![Built With AIWG](https://aiwg.io/assets/badges/built-with-aiwg-dark.png)](https://aiwg.io)
2016
2515
 
@@ -2018,24 +2517,30 @@ Building on AIWG? Show it. Grab a **Built With AIWG** / **Powered By AIWG** badg
2018
2517
  [![Built With AIWG](https://aiwg.io/assets/badges/built-with-aiwg-dark.png)](https://aiwg.io)
2019
2518
  ```
2020
2519
 
2021
- Full set with copy-paste snippets at **[aiwg.io/badges](https://aiwg.io/badges)**.
2520
+ Full set with copy-paste snippets: **[aiwg.io/badges](https://aiwg.io/badges)**.
2022
2521
 
2023
2522
  ---
2024
2523
 
2025
2524
  ## Usage Notes
2026
2525
 
2027
- AIWG is optimized for token efficiency. Rules deploy as a consolidated index (~200 lines) instead of 35 individual files (~9,321 lines). Most users on **Claude Pro** or similar plans will have no issues. See [Usage Notes](docs/usage-notes.md) for rate limit guidance.
2526
+ AIWG tries to keep always-loaded context small by using kernel quickrefs,
2527
+ provider-facing indexes, and on-demand discovery. Actual token use depends on
2528
+ the provider, selected workflows, project size, and how much context the agent
2529
+ loads. See [Usage Notes](docs/usage-notes.md) for rate-limit guidance.
2028
2530
 
2029
2531
  ---
2030
2532
 
2031
2533
  ## License
2032
2534
 
2033
- AIWG-authored code is available under the **MIT License**. See [LICENSE](LICENSE).
2535
+ AIWG-authored code is available under the **MIT License**. See
2536
+ [LICENSE](LICENSE).
2034
2537
  Runtime dependencies retain their own licenses; see
2035
2538
  [THIRD_PARTY_NOTICES.md](THIRD_PARTY_NOTICES.md) for the reviewed Fortemi and
2036
2539
  Bytecask AGPL boundary, source links, and inspection instructions.
2037
2540
 
2038
- **Important:** This framework does not provide legal, security, or financial advice. All generated content should be reviewed before use. See [Terms of Use](docs/terms.md) for full disclaimers.
2541
+ This framework does not provide legal, security, financial, medical, or other
2542
+ professional advice. Generated work should be reviewed before use. See
2543
+ [Terms of Use](docs/terms.md) for full terms.
2039
2544
 
2040
2545
  ---
2041
2546
 
@@ -2079,9 +2584,16 @@ Custom AI and blockchain solutions for the digital age.
2079
2584
 
2080
2585
  ## Acknowledgments
2081
2586
 
2082
- **Research foundations:** Built on established principles from cognitive science (Miller 1956, Sweller 1988), multi-agent systems (Jacobs et al. 1991, MetaGPT, AutoGen), software engineering (Cooper 1990, RUP), and recent AI systems research (ReAct, Self-Refine, DSPy, SWE-Agent). Implements standards from FAIR Principles, OAIS (ISO 14721), W3C PROV, GRADE evidence assessment, and MCP protocol (Linux Foundation).
2587
+ **Research foundations:** AIWG draws from cognitive science (Miller 1956,
2588
+ Sweller 1988), multi-agent systems (Jacobs et al. 1991, MetaGPT, AutoGen),
2589
+ software engineering (Cooper 1990, RUP), and AI systems research including
2590
+ ReAct, Self-Refine, DSPy, and SWE-agent. Standards and methods such as FAIR,
2591
+ OAIS, W3C PROV, GRADE, and MCP inform the structure of selected workflows and
2592
+ artifacts.
2083
2593
 
2084
- **Platforms:** Thanks to Anthropic (Claude Code), GitHub (Copilot), Warp, Factory AI, Cursor, and the OpenCode community for building the platforms that enable this work.
2594
+ **Platforms:** Thanks to Anthropic (Claude Code), GitHub (Copilot), Warp,
2595
+ Factory AI, Cursor, OpenCode, and other provider communities for building
2596
+ tools that make project-local AI workflows possible.
2085
2597
 
2086
2598
  ---
2087
2599