tribunal-kit 4.5.1 → 4.6.1

This diff represents the content of publicly available package versions that have been released to one of the supported registries. The information contained in this diff is provided for informational purposes only and reflects changes between package versions as they appear in their respective public registries.
Files changed (214) hide show
  1. package/.agent/.shared/ui-ux-pro-max/README.md +4 -4
  2. package/.agent/ARCHITECTURE.md +282 -277
  3. package/.agent/agents/accessibility-reviewer.md +187 -187
  4. package/.agent/agents/ai-code-reviewer.md +199 -199
  5. package/.agent/agents/api-architect.md +71 -66
  6. package/.agent/agents/backend-specialist.md +219 -215
  7. package/.agent/agents/cloud-engineer.md +98 -0
  8. package/.agent/agents/code-archaeologist.md +168 -161
  9. package/.agent/agents/database-architect.md +184 -184
  10. package/.agent/agents/db-latency-auditor.md +213 -216
  11. package/.agent/agents/debugger.md +198 -191
  12. package/.agent/agents/dependency-reviewer.md +106 -103
  13. package/.agent/agents/devops-engineer.md +218 -218
  14. package/.agent/agents/documentation-writer.md +209 -201
  15. package/.agent/agents/explorer-agent.md +167 -160
  16. package/.agent/agents/frontend-reviewer.md +162 -160
  17. package/.agent/agents/frontend-specialist.md +257 -248
  18. package/.agent/agents/game-developer.md +48 -48
  19. package/.agent/agents/logic-reviewer.md +118 -116
  20. package/.agent/agents/mobile-developer.md +197 -200
  21. package/.agent/agents/mobile-reviewer.md +159 -162
  22. package/.agent/agents/orchestrator.md +187 -181
  23. package/.agent/agents/penetration-tester.md +160 -157
  24. package/.agent/agents/performance-optimizer.md +183 -183
  25. package/.agent/agents/performance-reviewer.md +178 -178
  26. package/.agent/agents/precedence-reviewer.md +251 -250
  27. package/.agent/agents/product-manager.md +149 -142
  28. package/.agent/agents/product-owner.md +81 -80
  29. package/.agent/agents/project-planner.md +152 -142
  30. package/.agent/agents/qa-automation-engineer.md +216 -225
  31. package/.agent/agents/resilience-reviewer.md +88 -88
  32. package/.agent/agents/schema-reviewer.md +67 -67
  33. package/.agent/agents/security-auditor.md +180 -174
  34. package/.agent/agents/seo-specialist.md +188 -193
  35. package/.agent/agents/sql-reviewer.md +159 -161
  36. package/.agent/agents/supervisor-agent.md +173 -184
  37. package/.agent/agents/swarm-worker-contracts.md +170 -166
  38. package/.agent/agents/swarm-worker-registry.md +92 -92
  39. package/.agent/agents/system-architect.md +85 -0
  40. package/.agent/agents/test-coverage-reviewer.md +158 -160
  41. package/.agent/agents/test-engineer.md +118 -118
  42. package/.agent/agents/throughput-optimizer.md +291 -299
  43. package/.agent/agents/type-safety-reviewer.md +182 -175
  44. package/.agent/agents/ui-ux-auditor.md +300 -292
  45. package/.agent/agents/vitals-reviewer.md +223 -223
  46. package/.agent/mcp_config.json +37 -40
  47. package/.agent/patterns/generator.md +11 -9
  48. package/.agent/patterns/inversion.md +14 -12
  49. package/.agent/patterns/pipeline.md +11 -9
  50. package/.agent/patterns/reviewer.md +15 -13
  51. package/.agent/patterns/tool-wrapper.md +11 -9
  52. package/.agent/routing_index.json +714 -0
  53. package/.agent/rules/GEMINI.md +359 -352
  54. package/.agent/scripts/compile_router.py +112 -0
  55. package/.agent/scripts/migrate_skills_frontmatter.py +64 -0
  56. package/.agent/scripts/strengthen_skills.js +1 -1
  57. package/.agent/skills/advanced-rag-pipelines/SKILL.md +56 -0
  58. package/.agent/skills/agent-organizer/SKILL.md +156 -150
  59. package/.agent/skills/agentic-patterns/SKILL.md +313 -315
  60. package/.agent/skills/ai-prompt-injection-defense/SKILL.md +190 -184
  61. package/.agent/skills/api-patterns/SKILL.md +253 -247
  62. package/.agent/skills/api-security-auditor/SKILL.md +195 -193
  63. package/.agent/skills/app-builder/SKILL.md +573 -572
  64. package/.agent/skills/app-builder/templates/SKILL.md +108 -115
  65. package/.agent/skills/app-builder/templates/astro-static/TEMPLATE.md +76 -76
  66. package/.agent/skills/app-builder/templates/chrome-extension/TEMPLATE.md +92 -92
  67. package/.agent/skills/app-builder/templates/cli-tool/TEMPLATE.md +88 -88
  68. package/.agent/skills/app-builder/templates/electron-desktop/TEMPLATE.md +88 -88
  69. package/.agent/skills/app-builder/templates/express-api/TEMPLATE.md +83 -83
  70. package/.agent/skills/app-builder/templates/flutter-app/TEMPLATE.md +90 -90
  71. package/.agent/skills/app-builder/templates/monorepo-turborepo/TEMPLATE.md +90 -90
  72. package/.agent/skills/app-builder/templates/nextjs-fullstack/TEMPLATE.md +126 -122
  73. package/.agent/skills/app-builder/templates/nextjs-saas/TEMPLATE.md +127 -122
  74. package/.agent/skills/app-builder/templates/nextjs-static/TEMPLATE.md +172 -169
  75. package/.agent/skills/app-builder/templates/nuxt-app/TEMPLATE.md +139 -134
  76. package/.agent/skills/app-builder/templates/python-fastapi/TEMPLATE.md +83 -83
  77. package/.agent/skills/app-builder/templates/react-native-app/TEMPLATE.md +122 -119
  78. package/.agent/skills/appflow-wireframe/SKILL.md +146 -145
  79. package/.agent/skills/architecture/SKILL.md +226 -219
  80. package/.agent/skills/authentication-best-practices/SKILL.md +197 -189
  81. package/.agent/skills/backend-security-expert/SKILL.md +16 -2
  82. package/.agent/skills/bash-linux/SKILL.md +179 -179
  83. package/.agent/skills/behavioral-modes/SKILL.md +239 -223
  84. package/.agent/skills/brainstorming/SKILL.md +498 -486
  85. package/.agent/skills/browser-native-ai/SKILL.md +57 -4
  86. package/.agent/skills/building-native-ui/SKILL.md +202 -202
  87. package/.agent/skills/cicd-pro/SKILL.md +442 -0
  88. package/.agent/skills/clean-code/SKILL.md +400 -381
  89. package/.agent/skills/cloud-architect/SKILL.md +439 -0
  90. package/.agent/skills/code-review-checklist/SKILL.md +203 -194
  91. package/.agent/skills/config-validator/SKILL.md +165 -165
  92. package/.agent/skills/containerization-pro/SKILL.md +452 -0
  93. package/.agent/skills/csharp-developer/SKILL.md +518 -518
  94. package/.agent/skills/data-validation-schemas/SKILL.md +333 -328
  95. package/.agent/skills/database-design/SKILL.md +247 -240
  96. package/.agent/skills/deployment-procedures/SKILL.md +172 -169
  97. package/.agent/skills/devops-engineer/SKILL.md +345 -345
  98. package/.agent/skills/devops-incident-responder/SKILL.md +143 -137
  99. package/.agent/skills/documentation-templates/SKILL.md +291 -279
  100. package/.agent/skills/edge-computing/SKILL.md +183 -181
  101. package/.agent/skills/emil-design-eng/SKILL.md +147 -0
  102. package/.agent/skills/error-resilience/SKILL.md +411 -428
  103. package/.agent/skills/extract-design-system/SKILL.md +160 -158
  104. package/.agent/skills/framer-motion-expert/SKILL.md +253 -244
  105. package/.agent/skills/frontend-design/SKILL.md +208 -201
  106. package/.agent/skills/frontend-security-expert/SKILL.md +16 -3
  107. package/.agent/skills/game-design-expert/SKILL.md +132 -129
  108. package/.agent/skills/game-engineering-expert/SKILL.md +148 -146
  109. package/.agent/skills/generative-ui-expert/SKILL.md +57 -1
  110. package/.agent/skills/geo-fundamentals/SKILL.md +148 -147
  111. package/.agent/skills/git-pro/SKILL.md +435 -0
  112. package/.agent/skills/github-operations/SKILL.md +335 -329
  113. package/.agent/skills/gsap-core/SKILL.md +319 -308
  114. package/.agent/skills/gsap-frameworks/SKILL.md +213 -207
  115. package/.agent/skills/gsap-performance/SKILL.md +139 -133
  116. package/.agent/skills/gsap-plugins/SKILL.md +486 -480
  117. package/.agent/skills/gsap-react/SKILL.md +202 -189
  118. package/.agent/skills/gsap-scrolltrigger/SKILL.md +357 -350
  119. package/.agent/skills/gsap-timeline/SKILL.md +165 -161
  120. package/.agent/skills/gsap-utils/SKILL.md +344 -338
  121. package/.agent/skills/harness-protocol/SKILL.md +48 -0
  122. package/.agent/skills/i18n-localization/SKILL.md +174 -163
  123. package/.agent/skills/intelligent-routing/SKILL.md +202 -246
  124. package/.agent/skills/knowledge-graph/SKILL.md +60 -52
  125. package/.agent/skills/lint-and-validate/SKILL.md +261 -261
  126. package/.agent/skills/llm-engineering/SKILL.md +400 -394
  127. package/.agent/skills/local-first/SKILL.md +178 -178
  128. package/.agent/skills/mcp-builder/SKILL.md +143 -142
  129. package/.agent/skills/mobile-design/SKILL.md +272 -263
  130. package/.agent/skills/monorepo-management/SKILL.md +335 -334
  131. package/.agent/skills/motion-engineering/SKILL.md +266 -234
  132. package/.agent/skills/nextjs-react-expert/SKILL.md +236 -234
  133. package/.agent/skills/nodejs-best-practices/SKILL.md +547 -548
  134. package/.agent/skills/observability/SKILL.md +343 -343
  135. package/.agent/skills/parallel-agents/SKILL.md +143 -146
  136. package/.agent/skills/performance-profiling/SKILL.md +259 -267
  137. package/.agent/skills/plan-writing/SKILL.md +150 -142
  138. package/.agent/skills/platform-engineer/SKILL.md +148 -147
  139. package/.agent/skills/playwright-best-practices/SKILL.md +188 -187
  140. package/.agent/skills/powershell-windows/SKILL.md +162 -162
  141. package/.agent/skills/project-idioms/SKILL.md +137 -137
  142. package/.agent/skills/python-patterns/SKILL.md +260 -259
  143. package/.agent/skills/python-pro/SKILL.md +324 -323
  144. package/.agent/skills/react-specialist/SKILL.md +305 -277
  145. package/.agent/skills/readme-builder/SKILL.md +310 -300
  146. package/.agent/skills/realtime-patterns/SKILL.md +323 -319
  147. package/.agent/skills/red-team-tactics/SKILL.md +231 -218
  148. package/.agent/skills/review-animations/SKILL.md +72 -0
  149. package/.agent/skills/review-animations/STANDARDS.md +73 -0
  150. package/.agent/skills/rust-pro/SKILL.md +671 -673
  151. package/.agent/skills/seo-fundamentals/SKILL.md +179 -179
  152. package/.agent/skills/server-management/SKILL.md +218 -214
  153. package/.agent/skills/shadcn-ui-expert/SKILL.md +231 -231
  154. package/.agent/skills/skill-creator/SKILL.md +87 -86
  155. package/.agent/skills/sql-pro/SKILL.md +629 -629
  156. package/.agent/skills/supabase-postgres-best-practices/SKILL.md +97 -97
  157. package/.agent/skills/swiftui-expert/SKILL.md +204 -201
  158. package/.agent/skills/system-design-pro/SKILL.md +345 -0
  159. package/.agent/skills/systematic-debugging/SKILL.md +153 -142
  160. package/.agent/skills/tailwind-patterns/SKILL.md +610 -566
  161. package/.agent/skills/tdd-workflow/SKILL.md +169 -161
  162. package/.agent/skills/test-result-analyzer/SKILL.md +313 -309
  163. package/.agent/skills/testing-patterns/SKILL.md +566 -579
  164. package/.agent/skills/trend-researcher/SKILL.md +243 -237
  165. package/.agent/skills/typescript-advanced/SKILL.md +336 -335
  166. package/.agent/skills/ui-ux-pro-max/SKILL.md +590 -562
  167. package/.agent/skills/ui-ux-researcher/SKILL.md +244 -244
  168. package/.agent/skills/vue-expert/SKILL.md +294 -275
  169. package/.agent/skills/vulnerability-scanner/SKILL.md +416 -404
  170. package/.agent/skills/web-accessibility-auditor/SKILL.md +219 -218
  171. package/.agent/skills/web-design-guidelines/SKILL.md +192 -186
  172. package/.agent/skills/webapp-testing/SKILL.md +167 -169
  173. package/.agent/skills/webgpu-performance/SKILL.md +56 -2
  174. package/.agent/skills/whimsy-injector/SKILL.md +346 -325
  175. package/.agent/skills/workflow-optimizer/SKILL.md +231 -229
  176. package/.agent/workflows/acf.md +141 -0
  177. package/.agent/workflows/api-tester.md +176 -151
  178. package/.agent/workflows/audit.md +150 -127
  179. package/.agent/workflows/brainstorm.md +134 -110
  180. package/.agent/workflows/changelog.md +140 -112
  181. package/.agent/workflows/create.md +168 -124
  182. package/.agent/workflows/debug.md +190 -165
  183. package/.agent/workflows/deploy.md +201 -180
  184. package/.agent/workflows/enhance.md +154 -128
  185. package/.agent/workflows/fix.md +136 -114
  186. package/.agent/workflows/generate.md +198 -183
  187. package/.agent/workflows/marathon.md +37 -11
  188. package/.agent/workflows/migrate.md +184 -160
  189. package/.agent/workflows/orchestrate.md +192 -168
  190. package/.agent/workflows/performance-benchmarker.md +135 -114
  191. package/.agent/workflows/plan.md +196 -173
  192. package/.agent/workflows/preview.md +103 -80
  193. package/.agent/workflows/refactor.md +192 -161
  194. package/.agent/workflows/review-ai.md +125 -101
  195. package/.agent/workflows/review.md +141 -116
  196. package/.agent/workflows/session.md +122 -94
  197. package/.agent/workflows/status.md +101 -79
  198. package/.agent/workflows/strengthen-skills.md +164 -138
  199. package/.agent/workflows/super-prompt.md +24 -0
  200. package/.agent/workflows/swarm.md +193 -179
  201. package/.agent/workflows/test.md +211 -189
  202. package/.agent/workflows/tribunal-backend.md +136 -105
  203. package/.agent/workflows/tribunal-database.md +129 -95
  204. package/.agent/workflows/tribunal-frontend.md +140 -96
  205. package/.agent/workflows/tribunal-full.md +131 -100
  206. package/.agent/workflows/tribunal-mobile.md +129 -95
  207. package/.agent/workflows/tribunal-performance.md +136 -110
  208. package/.agent/workflows/tribunal-speed.md +209 -183
  209. package/.agent/workflows/ui-ux-pro-max.md +155 -122
  210. package/README.md +107 -55
  211. package/mcp_config.json +1 -3
  212. package/package.json +94 -94
  213. package/.agent/GEMINI.md +0 -121
  214. package/.agent/skills/doc.md +0 -177
@@ -1,183 +1,198 @@
1
- ---
2
- description: Generate code using the full Tribunal Anti-Hallucination pipeline. Maker generates grounded in real project context at low temperature → domain-selected reviewers audit in parallel → Human Gate for final approval. Nothing is written to disk without explicit approval.
3
- ---
4
-
5
- # /generate — Hallucination-Free Code Generation
6
-
7
- $ARGUMENTS
8
-
9
- ---
10
-
11
- ## When to Use /generate
12
-
13
- | Use `/generate` when... | Use something else when... |
14
- |:---|:---|
15
- | New code needs to be written from scratch | Existing code needs modification `/enhance` |
16
- | A single focused piece of code is needed | Multi-domain build → `/create` or `/swarm` |
17
- | A safe, reviewed snippet is required | You want to understand options first → `/plan` |
18
- | You need a quick but Tribunal-reviewed piece | Full project structure needed → `/create` |
19
-
20
- ---
21
-
22
- ## Pipeline Flow
23
-
24
- ```
25
- Your request
26
-
27
-
28
- [Phase 6] Context Broker — Skill Selection
29
- ├── Scores all 90+ skills against your task keywords
30
- ├── Level 0 (Essential): top skills full content, injected first
31
- ├── Level 1 (Supplementary): medium relevance key rules only
32
- ├── Level 2 (Available): listed for reference only
33
- └── Large models: Essential + Supplementary | Small models: Essential only
34
-
35
-
36
- Context scan (MANDATORY before first line of code)
37
- ├── Read package.json → verify all imports exist
38
- ├── Read tsconfig.json → understand strictness, paths aliases
39
- ├── Read referenced files → understand actual data shapes
40
- └── Read .env.example → know available environment variables
41
-
42
-
43
- Maker generates at temperature 0.1
44
- ├── Only methods verified in official docs
45
- ├── Only packages in package.json
46
- ├── // VERIFY: [reason] on any uncertain call
47
- └── No full application generation modules only
48
-
49
-
50
- [Phase 6] Inner-Loop Validator (AUTO — runs before you see the code)
51
- ├── Scans generated snippet for OWASP patterns (critical/high/medium/low)
52
- ├── Runs structural heuristics (empty catch, throw strings, env without fallback)
53
- ├── Verdict: APPROVED / WARNING / REJECTED
54
- │ ├── APPROVEDcontinues to Tribunal Review
55
- ├── WARNING → noted, continues with flag
56
- │ └── REJECTED → Maker auto-corrects (up to 2 inner-loop attempts)
57
- └── Only clean code reaches the Tribunal reviewers
58
-
59
-
60
- Tribunal Reviewers run in parallel (auto-selected by keyword)
61
-
62
-
63
- Human Gate — verdicts shown + unified diff
64
- Y = write to disk | N = discard | R = revise with feedback
65
- ```
66
-
67
- ---
68
-
69
- ## What the Maker Is Not Allowed to Do
70
-
71
- ```
72
- ❌ Import a package not in package.json
73
- ❌ Call a method not verified in official documentation
74
- Use TypeScript 'any' without an explanation comment
75
- ❌ Generate an entire application in one shot
76
- ❌ Guess at database column or table names (read schema first)
77
- Fabricate API response shapes (read existing types first)
78
- Assume environment variables exist (read .env.example first)
79
- ❌ Use Next.js 14 patterns in a Next.js 15 project (check version!)
80
- ❌ Use React 18 hooks in a React 19 project (useFormState → useActionState)
81
- ❌ Use framer-motion v6 API in a v12 project (exitBeforeEnter → mode="wait")
82
- ❌ Use raw useEffect for GSAP — always useGSAP from @gsap/react
83
- Hallucinate LLM model names verify against provider's current model list
84
- ```
85
-
86
- When unsure: write `// VERIFY: [specific reason]` instead of hallucinating.
87
-
88
- ---
89
-
90
- ## Reviewer Auto-Selection
91
-
92
- **Always active:**
93
- ```
94
- precedence-reviewer→ Enforces repository Case Law and past rejections (Runs First)
95
- logic-reviewer → Hallucinated methods, undefined refs, impossible logic
96
- security-auditor → OWASP vulnerabilities, hardcoded secrets, injection
97
- ```
98
-
99
- **Auto-activated by keywords:**
100
-
101
- | Keyword in request | Additional Reviewers |
102
- |:---|:---|
103
- | `api`, `route`, `endpoint`, `handler`, `server action` | `dependency-reviewer` + `type-safety-reviewer` |
104
- | `sql`, `query`, `database`, `prisma`, `drizzle`, `orm` | `sql-reviewer` |
105
- | `component`, `hook`, `react`, `vue`, `jsx`, `tsx` | `frontend-reviewer` + `type-safety-reviewer` + `ui-ux-auditor` |
106
- | `ui`, `design`, `landing`, `page`, `layout`, `style`, `css` | `ui-ux-auditor` + `accessibility-reviewer` |
107
- | `animation`, `gsap`, `framer`, `motion`, `scroll` | `frontend-reviewer` + `performance-reviewer` + `ui-ux-auditor` |
108
- | `test`, `spec`, `vitest`, `jest`, `playwright` | `test-coverage-reviewer` |
109
- | `slow`, `optimize`, `cache`, `performance`, `bundle` | `performance-reviewer` |
110
- | `mobile`, `react native`, `expo` | `mobile-reviewer` |
111
- | `llm`, `openai`, `anthropic`, `gemini`, `embedding`, `ai` | `ai-code-reviewer` |
112
- | `aria`, `wcag`, `a11y`, `accessibility` | `accessibility-reviewer` + `ui-ux-auditor` |
113
- | `import`, `package`, `npm`, `require` | `dependency-reviewer` |
114
-
115
- > For maximum safety on critical code: use `/tribunal-full` for all 11 reviewers simultaneously.
116
-
117
- ---
118
-
119
- ## Reviewer Verdicts
120
-
121
- | Verdict | Meaning | What Happens |
122
- |:---|:---|:---|
123
- | `✅ APPROVED` | No issues found | Proceeds to Human Gate |
124
- | `⚠️ WARNING` | Non-blocking issue | Human Gate shown with warning highlighted |
125
- | `❌ REJECTED` | Blocking issue | Maker revises before Human Gate |
126
-
127
- **Retry limit:** Maker is revised up to 3 times per REJECTED verdict. After 3 failures, the session halts and reports to the user with full failure history. No silent failures.
128
-
129
- ---
130
-
131
- ## Output Format
132
-
133
- ```
134
- ━━━ Tribunal: [Domain] ━━━━━━━━━━━━━━━━━━━━━━
135
-
136
- Active reviewers: logic · security · [others]
137
-
138
- [Generated code with // VERIFY: tags where uncertain]
139
-
140
- ━━━ Verdicts ━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━
141
-
142
- logic-reviewer: ✅ APPROVED
143
- security-auditor: ✅ APPROVED
144
- dependency-reviewer: ⚠️ WARNING — lodash not in package.json
145
-
146
- ━━━ Warnings ━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━
147
-
148
- dependency-reviewer:
149
- ⚠️ Medium — Line 3: 'lodash' imported but not in package.json
150
- Fix: npm install lodash OR use built-in Array methods
151
-
152
- ━━━ Human Gate ━━━━━━━━━━━━━━━━━━━━━━━━━━━━
153
- Write to disk? Y = approve | N = discard | R = revise with feedback
154
- ```
155
-
156
- ---
157
-
158
- | After /generate shows... | Go to |
159
- |:---|:---|
160
- | Multiple files need changing | `/enhance` for impact-zone analysis |
161
- | Security-critical code was generated | `/tribunal-full` for maximum coverage |
162
- | DB queries were generated | `/tribunal-database` |
163
- | New API routes were generated | `/tribunal-backend` |
164
- | Animation/motion code generated | `/tribunal-frontend` |
165
- | Tests need to be written next | `/test` |
166
- | Something was rejected 3 times | Escalate to human with failure report |
167
-
168
- ---
169
-
170
- ## Usage Examples
171
-
172
- ```
173
- /generate a JWT middleware for Express with HS256 algorithm enforcement
174
- /generate a Prisma query for users with their published posts in last 30 days
175
- /generate a debounced search hook in React 19 using useDeferredValue
176
- /generate a parameterized SQL query for paginated order history
177
- /generate a Zod schema for email + password + role login input
178
- /generate a Server Action for creating a product with image upload
179
- /generate a rate-limited fetch wrapper using @upstash/ratelimit
180
- /generate a Framer Motion page transition with shared element (layoutId)
181
- /generate a GSAP ScrollTrigger timeline with useGSAP React hook
182
- /generate an OpenAI structured output call with Zod schema validation
183
- ```
1
+ ---
2
+ description: Generate code using the full Tribunal Anti-Hallucination pipeline. Maker generates grounded in real project context at low temperature → domain-selected reviewers audit in parallel → Human Gate for final approval. Nothing is written to disk without explicit approval.
3
+ required-skills: auto-detected from request keywords (see Reviewer Auto-Selection table)
4
+ ---
5
+
6
+ # /generate — Hallucination-Free Code Generation
7
+
8
+ $ARGUMENTS
9
+
10
+ ---
11
+
12
+ ## $CONTEXT_REQUIRED
13
+
14
+ ```
15
+ Read BEFORE generating any code (MANDATORY never skip):
16
+ package.json → Verify all imports exist before using them
17
+ tsconfig.json → Understand strictness, paths aliases, target version
18
+ context/*.acf → Load explicit goals, rules, and boundaries
19
+ □ .env.example → Know available environment variables
20
+ □ Referenced files → Understand actual data shapes and types
21
+ ```
22
+
23
+ ---
24
+
25
+ ## When to Use /generate
26
+
27
+ | Use `/generate` when... | Use something else when... |
28
+ | :------------------------------------------- | :--------------------------------------------- |
29
+ | New code needs to be written from scratch | Existing code needs modification → `/enhance` |
30
+ | A single focused piece of code is needed | Multi-domain build → `/create` or `/swarm` |
31
+ | A safe, reviewed snippet is required | You want to understand options first → `/plan` |
32
+ | You need a quick but Tribunal-reviewed piece | Full project structure needed → `/create` |
33
+
34
+ ---
35
+
36
+ ## Pipeline Flow
37
+
38
+ ```
39
+ Your request
40
+
41
+
42
+ [Phase 6] Context Broker — Skill Selection
43
+ ├── Scores all 90+ skills against your task keywords
44
+ ├── Level 0 (Essential): top skills — full content, injected first
45
+ ├── Level 1 (Supplementary): medium relevance — key rules only
46
+ ├── Level 2 (Available): listed for reference only
47
+ └── Large models: Essential + Supplementary | Small models: Essential only
48
+
49
+
50
+ Context scan (MANDATORY before first line of code)
51
+ ├── Read package.json verify all imports exist
52
+ ├── Read tsconfig.json understand strictness, paths aliases
53
+ ├── Read referenced files understand actual data shapes
54
+ └── Read .env.example know available environment variables
55
+
56
+
57
+ Maker generates at temperature 0.1
58
+ ├── Only methods verified in official docs
59
+ ├── Only packages in package.json
60
+ ├── // VERIFY: [reason] on any uncertain call
61
+ └── No full application generation — modules only
62
+
63
+
64
+ [Phase 6] Inner-Loop Validator (AUTO runs before you see the code)
65
+ ├── Scans generated snippet for OWASP patterns (critical/high/medium/low)
66
+ ├── Runs structural heuristics (empty catch, throw strings, env without fallback)
67
+ ├── Verdict: APPROVED / WARNING / REJECTED
68
+ │ ├── APPROVED → continues to Tribunal Review
69
+ │ ├── WARNING → noted, continues with flag
70
+ │ └── REJECTED → Maker auto-corrects (up to 2 inner-loop attempts)
71
+ └── Only clean code reaches the Tribunal reviewers
72
+
73
+
74
+ Tribunal Reviewers run in parallel (auto-selected by keyword)
75
+
76
+
77
+ Human Gate verdicts shown + unified diff
78
+ Y = write to disk | N = discard | R = revise with feedback
79
+ ```
80
+
81
+ ---
82
+
83
+ ## What the Maker Is Not Allowed to Do
84
+
85
+ ```
86
+ Import a package not in package.json
87
+ ❌ Call a method not verified in official documentation
88
+ ❌ Use TypeScript 'any' without an explanation comment
89
+ ❌ Generate an entire application in one shot
90
+ Guess at database column or table names (read schema first)
91
+ ❌ Fabricate API response shapes (read existing types first)
92
+ Assume environment variables exist (read .env.example first)
93
+ ❌ Use Next.js 14 patterns in a Next.js 15 project (check version!)
94
+ Use React 18 hooks in a React 19 project (useFormState → useActionState)
95
+ ❌ Use framer-motion v6 API in a v12 project (exitBeforeEnter → mode="wait")
96
+ Use raw useEffect for GSAP — always useGSAP from @gsap/react
97
+ ❌ Hallucinate LLM model names — verify against provider's current model list
98
+ ```
99
+
100
+ When unsure: write `// VERIFY: [specific reason]` instead of hallucinating.
101
+
102
+ ---
103
+
104
+ ## Reviewer Auto-Selection
105
+
106
+ **Always active:**
107
+
108
+ ```
109
+ precedence-reviewer→ Enforces repository Case Law and past rejections (Runs First)
110
+ logic-reviewer → Hallucinated methods, undefined refs, impossible logic
111
+ security-auditor → OWASP vulnerabilities, hardcoded secrets, injection
112
+ ```
113
+
114
+ **Auto-activated by keywords:**
115
+
116
+ | Keyword in request | Additional Reviewers |
117
+ | :---------------------------------------------------------- | :------------------------------------------------------------- |
118
+ | `api`, `route`, `endpoint`, `handler`, `server action` | `dependency-reviewer` + `type-safety-reviewer` |
119
+ | `sql`, `query`, `database`, `prisma`, `drizzle`, `orm` | `sql-reviewer` |
120
+ | `component`, `hook`, `react`, `vue`, `jsx`, `tsx` | `frontend-reviewer` + `type-safety-reviewer` + `ui-ux-auditor` |
121
+ | `ui`, `design`, `landing`, `page`, `layout`, `style`, `css` | `ui-ux-auditor` + `accessibility-reviewer` |
122
+ | `animation`, `gsap`, `framer`, `motion`, `scroll` | `frontend-reviewer` + `performance-reviewer` + `ui-ux-auditor` |
123
+ | `test`, `spec`, `vitest`, `jest`, `playwright` | `test-coverage-reviewer` |
124
+ | `slow`, `optimize`, `cache`, `performance`, `bundle` | `performance-reviewer` |
125
+ | `mobile`, `react native`, `expo` | `mobile-reviewer` |
126
+ | `llm`, `openai`, `anthropic`, `gemini`, `embedding`, `ai` | `ai-code-reviewer` |
127
+ | `aria`, `wcag`, `a11y`, `accessibility` | `accessibility-reviewer` + `ui-ux-auditor` |
128
+ | `import`, `package`, `npm`, `require` | `dependency-reviewer` |
129
+
130
+ > For maximum safety on critical code: use `/tribunal-full` for all 18 reviewers simultaneously.
131
+
132
+ ---
133
+
134
+ ## Reviewer Verdicts
135
+
136
+ | Verdict | Meaning | What Happens |
137
+ | :------------ | :----------------- | :---------------------------------------- |
138
+ | `✅ APPROVED` | No issues found | Proceeds to Human Gate |
139
+ | `⚠️ WARNING` | Non-blocking issue | Human Gate shown with warning highlighted |
140
+ | `❌ REJECTED` | Blocking issue | Maker revises before Human Gate |
141
+
142
+ **Retry limit:** Maker is revised up to 3 times per REJECTED verdict. After 3 failures, the session halts and reports to the user with full failure history. No silent failures.
143
+
144
+ ---
145
+
146
+ ## Output Format
147
+
148
+ ```
149
+ ━━━ Tribunal: [Domain] ━━━━━━━━━━━━━━━━━━━━━━
150
+
151
+ Active reviewers: logic · security · [others]
152
+
153
+ [Generated code with // VERIFY: tags where uncertain]
154
+
155
+ ━━━ Verdicts ━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━
156
+
157
+ logic-reviewer: ✅ APPROVED
158
+ security-auditor: ✅ APPROVED
159
+ dependency-reviewer: ⚠️ WARNING — lodash not in package.json
160
+
161
+ ━━━ Warnings ━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━
162
+
163
+ dependency-reviewer:
164
+ ⚠️ Medium Line 3: 'lodash' imported but not in package.json
165
+ Fix: npm install lodash OR use built-in Array methods
166
+
167
+ ━━━ Human Gate ━━━━━━━━━━━━━━━━━━━━━━━━━━━━
168
+ Write to disk? Y = approve | N = discard | R = revise with feedback
169
+ ```
170
+
171
+ ---
172
+
173
+ | After /generate shows... | Go to |
174
+ | :----------------------------------- | :------------------------------------ |
175
+ | Multiple files need changing | `/enhance` for impact-zone analysis |
176
+ | Security-critical code was generated | `/tribunal-full` for maximum coverage |
177
+ | DB queries were generated | `/tribunal-database` |
178
+ | New API routes were generated | `/tribunal-backend` |
179
+ | Animation/motion code generated | `/tribunal-frontend` |
180
+ | Tests need to be written next | `/test` |
181
+ | Something was rejected 3 times | Escalate to human with failure report |
182
+
183
+ ---
184
+
185
+ ## Usage Examples
186
+
187
+ ```
188
+ /generate a JWT middleware for Express with HS256 algorithm enforcement
189
+ /generate a Prisma query for users with their published posts in last 30 days
190
+ /generate a debounced search hook in React 19 using useDeferredValue
191
+ /generate a parameterized SQL query for paginated order history
192
+ /generate a Zod schema for email + password + role login input
193
+ /generate a Server Action for creating a product with image upload
194
+ /generate a rate-limited fetch wrapper using @upstash/ratelimit
195
+ /generate a Framer Motion page transition with shared element (layoutId)
196
+ /generate a GSAP ScrollTrigger timeline with useGSAP React hook
197
+ /generate an OpenAI structured output call with Zod schema validation
198
+ ```
@@ -1,5 +1,6 @@
1
1
  ---
2
2
  description: Long-running agent harness for multi-session projects. Decomposes specs into atomic features tracked in JSON, ensures clean handoffs between sessions, and provides structured progress tracking. Based on Anthropic's long-running agent patterns.
3
+ required-skills: harness-protocol, agent-organizer
3
4
  ---
4
5
 
5
6
  # /marathon — Long-Running Agent Harness
@@ -8,14 +9,25 @@ $ARGUMENTS
8
9
 
9
10
  ---
10
11
 
12
+ ## $CONTEXT_REQUIRED
13
+
14
+ ```
15
+ Read BEFORE marathon start/continue:
16
+ □ progress.json → See current marathon state
17
+ □ feature_list.json → View remaining tasks
18
+ □ git log --oneline -5 → Check recent commits
19
+ ```
20
+
21
+ ---
22
+
11
23
  ## When to Use /marathon
12
24
 
13
- |Use `/marathon` when...|Use something else when...|
14
- |:---|:---|
15
- |A project requires multiple sessions to complete|Quick one-shot task → `/generate`|
16
- |You need structured progress tracking across context windows|Single feature addition → `/enhance`|
17
- |Building a complex app from a high-level spec|Planning without execution → `/plan`|
18
- |Previous agent sessions lost context or declared victory too early|Brainstorming options → `/brainstorm`|
25
+ | Use `/marathon` when... | Use something else when... |
26
+ | :----------------------------------------------------------------- | :------------------------------------ |
27
+ | A project requires multiple sessions to complete | Quick one-shot task → `/generate` |
28
+ | You need structured progress tracking across context windows | Single feature addition → `/enhance` |
29
+ | Building a complex app from a high-level spec | Planning without execution → `/plan` |
30
+ | Previous agent sessions lost context or declared victory too early | Brainstorming options → `/brainstorm` |
19
31
 
20
32
  ---
21
33
 
@@ -66,11 +78,7 @@ Features are stored in `.agent/history/marathon/feature_list.json` as structured
66
78
  "id": 1,
67
79
  "category": "core",
68
80
  "description": "User can open a new chat and see a welcome screen",
69
- "steps": [
70
- "Navigate to main page",
71
- "Click 'New Chat' button",
72
- "Verify welcome state renders"
73
- ],
81
+ "steps": ["Navigate to main page", "Click 'New Chat' button", "Verify welcome state renders"],
74
82
  "passes": false,
75
83
  "sessionCompleted": null
76
84
  }
@@ -90,9 +98,11 @@ Every new session starts by orienting the agent. This is the **Coding Agent** pa
90
98
  ### Steps
91
99
 
92
100
  1. **Read marathon state:**
101
+
93
102
  ```bash
94
103
  node .agent/scripts/marathon_harness.js session-start
95
104
  ```
105
+
96
106
  This automatically:
97
107
  - Reads `progress.json` and shows what was done in previous sessions
98
108
  - Reads `git log --oneline -20` for recent commits
@@ -100,6 +110,7 @@ Every new session starts by orienting the agent. This is the **Coding Agent** pa
100
110
  - Records the session start time
101
111
 
102
112
  2. **Start the dev environment:**
113
+
103
114
  ```bash
104
115
  node .agent/scripts/auto_preview.js start
105
116
  ```
@@ -147,6 +158,7 @@ Work on exactly **one feature at a time**. This incremental approach prevents th
147
158
  ### If a feature cannot be completed
148
159
 
149
160
  If a feature is blocked or too complex for the current session:
161
+
150
162
  1. Leave it as `passes: false`
151
163
  2. Add a log note explaining why:
152
164
  ```bash
@@ -169,11 +181,13 @@ Every session MUST leave the codebase in a clean, merge-ready state.
169
181
  - If something is half-done, either complete it or revert it
170
182
 
171
183
  2. **Record session end:**
184
+
172
185
  ```bash
173
186
  node .agent/scripts/marathon_harness.js session-end "Implemented dark mode, user settings page, and notification bell"
174
187
  ```
175
188
 
176
189
  3. **Final git commit:**
190
+
177
191
  ```bash
178
192
  git commit -m "marathon: session N complete, 15/47 features passing"
179
193
  ```
@@ -245,3 +259,15 @@ node .agent/scripts/marathon_harness.js reset
245
259
  /marathon status
246
260
  /marathon reset
247
261
  ```
262
+
263
+ ---
264
+
265
+ ## After /marathon — Next Steps
266
+
267
+ | Outcome | Next Command |
268
+ | :------------------------ | :------------------------------------ |
269
+ | Session completes | → `/marathon continue` (next session) |
270
+ | Feature done, needs audit | → `/audit` for project health |
271
+ | Marathon fully completed | → `/deploy` to ship |
272
+
273
+ ---