tribunal-kit 4.5.0 → 4.6.0

This diff represents the content of publicly available package versions that have been released to one of the supported registries. The information contained in this diff is provided for informational purposes only and reflects changes between package versions as they appear in their respective public registries.
Files changed (217) hide show
  1. package/.agent/.shared/ui-ux-pro-max/README.md +4 -4
  2. package/.agent/ARCHITECTURE.md +279 -277
  3. package/.agent/GEMINI.md +127 -121
  4. package/.agent/agents/accessibility-reviewer.md +187 -187
  5. package/.agent/agents/ai-code-reviewer.md +199 -199
  6. package/.agent/agents/api-architect.md +71 -66
  7. package/.agent/agents/backend-specialist.md +219 -215
  8. package/.agent/agents/cloud-engineer.md +98 -0
  9. package/.agent/agents/code-archaeologist.md +168 -161
  10. package/.agent/agents/database-architect.md +184 -184
  11. package/.agent/agents/db-latency-auditor.md +213 -216
  12. package/.agent/agents/debugger.md +198 -191
  13. package/.agent/agents/dependency-reviewer.md +106 -103
  14. package/.agent/agents/devops-engineer.md +218 -218
  15. package/.agent/agents/documentation-writer.md +209 -201
  16. package/.agent/agents/explorer-agent.md +167 -160
  17. package/.agent/agents/frontend-reviewer.md +162 -160
  18. package/.agent/agents/frontend-specialist.md +257 -248
  19. package/.agent/agents/game-developer.md +48 -48
  20. package/.agent/agents/logic-reviewer.md +118 -116
  21. package/.agent/agents/mobile-developer.md +197 -200
  22. package/.agent/agents/mobile-reviewer.md +159 -162
  23. package/.agent/agents/orchestrator.md +187 -181
  24. package/.agent/agents/penetration-tester.md +160 -157
  25. package/.agent/agents/performance-optimizer.md +183 -183
  26. package/.agent/agents/performance-reviewer.md +178 -178
  27. package/.agent/agents/precedence-reviewer.md +251 -250
  28. package/.agent/agents/product-manager.md +149 -142
  29. package/.agent/agents/product-owner.md +81 -80
  30. package/.agent/agents/project-planner.md +152 -142
  31. package/.agent/agents/qa-automation-engineer.md +216 -225
  32. package/.agent/agents/resilience-reviewer.md +88 -88
  33. package/.agent/agents/schema-reviewer.md +67 -67
  34. package/.agent/agents/security-auditor.md +180 -174
  35. package/.agent/agents/seo-specialist.md +188 -193
  36. package/.agent/agents/sql-reviewer.md +159 -161
  37. package/.agent/agents/supervisor-agent.md +173 -184
  38. package/.agent/agents/swarm-worker-contracts.md +170 -166
  39. package/.agent/agents/swarm-worker-registry.md +92 -92
  40. package/.agent/agents/system-architect.md +85 -0
  41. package/.agent/agents/test-coverage-reviewer.md +158 -160
  42. package/.agent/agents/test-engineer.md +118 -118
  43. package/.agent/agents/throughput-optimizer.md +291 -299
  44. package/.agent/agents/type-safety-reviewer.md +182 -175
  45. package/.agent/agents/ui-ux-auditor.md +300 -292
  46. package/.agent/agents/vitals-reviewer.md +223 -223
  47. package/.agent/mcp_config.json +37 -40
  48. package/.agent/patterns/generator.md +11 -9
  49. package/.agent/patterns/inversion.md +14 -12
  50. package/.agent/patterns/pipeline.md +11 -9
  51. package/.agent/patterns/reviewer.md +15 -13
  52. package/.agent/patterns/tool-wrapper.md +11 -9
  53. package/.agent/routing_index.json +654 -0
  54. package/.agent/rules/GEMINI.md +358 -352
  55. package/.agent/scripts/compile_router.py +112 -0
  56. package/.agent/scripts/migrate_skills_frontmatter.py +64 -0
  57. package/.agent/scripts/strengthen_skills.js +1 -1
  58. package/.agent/skills/advanced-rag-pipelines/SKILL.md +56 -0
  59. package/.agent/skills/agent-organizer/SKILL.md +156 -150
  60. package/.agent/skills/agentic-patterns/SKILL.md +313 -315
  61. package/.agent/skills/ai-prompt-injection-defense/SKILL.md +190 -184
  62. package/.agent/skills/api-patterns/SKILL.md +253 -247
  63. package/.agent/skills/api-security-auditor/SKILL.md +195 -193
  64. package/.agent/skills/app-builder/SKILL.md +573 -572
  65. package/.agent/skills/app-builder/templates/SKILL.md +108 -115
  66. package/.agent/skills/app-builder/templates/astro-static/TEMPLATE.md +76 -76
  67. package/.agent/skills/app-builder/templates/chrome-extension/TEMPLATE.md +92 -92
  68. package/.agent/skills/app-builder/templates/cli-tool/TEMPLATE.md +88 -88
  69. package/.agent/skills/app-builder/templates/electron-desktop/TEMPLATE.md +88 -88
  70. package/.agent/skills/app-builder/templates/express-api/TEMPLATE.md +83 -83
  71. package/.agent/skills/app-builder/templates/flutter-app/TEMPLATE.md +90 -90
  72. package/.agent/skills/app-builder/templates/monorepo-turborepo/TEMPLATE.md +90 -90
  73. package/.agent/skills/app-builder/templates/nextjs-fullstack/TEMPLATE.md +126 -122
  74. package/.agent/skills/app-builder/templates/nextjs-saas/TEMPLATE.md +127 -122
  75. package/.agent/skills/app-builder/templates/nextjs-static/TEMPLATE.md +172 -169
  76. package/.agent/skills/app-builder/templates/nuxt-app/TEMPLATE.md +139 -134
  77. package/.agent/skills/app-builder/templates/python-fastapi/TEMPLATE.md +83 -83
  78. package/.agent/skills/app-builder/templates/react-native-app/TEMPLATE.md +122 -119
  79. package/.agent/skills/appflow-wireframe/SKILL.md +146 -145
  80. package/.agent/skills/architecture/SKILL.md +226 -219
  81. package/.agent/skills/authentication-best-practices/SKILL.md +197 -189
  82. package/.agent/skills/backend-security-expert/SKILL.md +16 -2
  83. package/.agent/skills/bash-linux/SKILL.md +179 -179
  84. package/.agent/skills/behavioral-modes/SKILL.md +239 -223
  85. package/.agent/skills/brainstorming/SKILL.md +498 -486
  86. package/.agent/skills/browser-native-ai/SKILL.md +57 -4
  87. package/.agent/skills/building-native-ui/SKILL.md +202 -202
  88. package/.agent/skills/cicd-pro/SKILL.md +442 -0
  89. package/.agent/skills/clean-code/SKILL.md +400 -381
  90. package/.agent/skills/cloud-architect/SKILL.md +439 -0
  91. package/.agent/skills/code-review-checklist/SKILL.md +203 -194
  92. package/.agent/skills/config-validator/SKILL.md +165 -165
  93. package/.agent/skills/containerization-pro/SKILL.md +452 -0
  94. package/.agent/skills/csharp-developer/SKILL.md +518 -518
  95. package/.agent/skills/data-validation-schemas/SKILL.md +333 -328
  96. package/.agent/skills/database-design/SKILL.md +247 -240
  97. package/.agent/skills/deployment-procedures/SKILL.md +172 -169
  98. package/.agent/skills/devops-engineer/SKILL.md +345 -345
  99. package/.agent/skills/devops-incident-responder/SKILL.md +143 -137
  100. package/.agent/skills/doc.md +209 -177
  101. package/.agent/skills/documentation-templates/SKILL.md +291 -279
  102. package/.agent/skills/edge-computing/SKILL.md +183 -181
  103. package/.agent/skills/error-resilience/SKILL.md +411 -428
  104. package/.agent/skills/extract-design-system/SKILL.md +160 -158
  105. package/.agent/skills/framer-motion-expert/SKILL.md +253 -244
  106. package/.agent/skills/frontend-design/SKILL.md +208 -201
  107. package/.agent/skills/frontend-security-expert/SKILL.md +16 -3
  108. package/.agent/skills/game-design-expert/SKILL.md +132 -129
  109. package/.agent/skills/game-engineering-expert/SKILL.md +148 -146
  110. package/.agent/skills/generative-ui-expert/SKILL.md +57 -1
  111. package/.agent/skills/geo-fundamentals/SKILL.md +148 -147
  112. package/.agent/skills/git-pro/SKILL.md +435 -0
  113. package/.agent/skills/github-operations/SKILL.md +335 -329
  114. package/.agent/skills/gsap-core/SKILL.md +319 -308
  115. package/.agent/skills/gsap-frameworks/SKILL.md +213 -207
  116. package/.agent/skills/gsap-performance/SKILL.md +139 -133
  117. package/.agent/skills/gsap-plugins/SKILL.md +486 -480
  118. package/.agent/skills/gsap-react/SKILL.md +202 -189
  119. package/.agent/skills/gsap-scrolltrigger/SKILL.md +357 -350
  120. package/.agent/skills/gsap-timeline/SKILL.md +165 -161
  121. package/.agent/skills/gsap-utils/SKILL.md +344 -338
  122. package/.agent/skills/harness-protocol/SKILL.md +48 -0
  123. package/.agent/skills/i18n-localization/SKILL.md +174 -163
  124. package/.agent/skills/intelligent-routing/SKILL.md +202 -246
  125. package/.agent/skills/knowledge-graph/SKILL.md +60 -52
  126. package/.agent/skills/lint-and-validate/SKILL.md +261 -261
  127. package/.agent/skills/llm-engineering/SKILL.md +400 -394
  128. package/.agent/skills/local-first/SKILL.md +178 -178
  129. package/.agent/skills/mcp-builder/SKILL.md +143 -142
  130. package/.agent/skills/mobile-design/SKILL.md +272 -263
  131. package/.agent/skills/monorepo-management/SKILL.md +335 -334
  132. package/.agent/skills/motion-engineering/SKILL.md +266 -234
  133. package/.agent/skills/nextjs-react-expert/SKILL.md +236 -234
  134. package/.agent/skills/nodejs-best-practices/SKILL.md +547 -548
  135. package/.agent/skills/observability/SKILL.md +343 -343
  136. package/.agent/skills/parallel-agents/SKILL.md +143 -146
  137. package/.agent/skills/performance-profiling/SKILL.md +259 -267
  138. package/.agent/skills/plan-writing/SKILL.md +150 -142
  139. package/.agent/skills/platform-engineer/SKILL.md +148 -147
  140. package/.agent/skills/playwright-best-practices/SKILL.md +188 -187
  141. package/.agent/skills/powershell-windows/SKILL.md +162 -162
  142. package/.agent/skills/project-idioms/SKILL.md +137 -137
  143. package/.agent/skills/python-patterns/SKILL.md +260 -259
  144. package/.agent/skills/python-pro/SKILL.md +324 -323
  145. package/.agent/skills/react-specialist/SKILL.md +305 -277
  146. package/.agent/skills/readme-builder/SKILL.md +310 -300
  147. package/.agent/skills/realtime-patterns/SKILL.md +323 -319
  148. package/.agent/skills/red-team-tactics/SKILL.md +231 -218
  149. package/.agent/skills/rust-pro/SKILL.md +671 -673
  150. package/.agent/skills/seo-fundamentals/SKILL.md +179 -179
  151. package/.agent/skills/server-management/SKILL.md +218 -214
  152. package/.agent/skills/shadcn-ui-expert/SKILL.md +231 -231
  153. package/.agent/skills/skill-creator/SKILL.md +87 -86
  154. package/.agent/skills/sql-pro/SKILL.md +629 -629
  155. package/.agent/skills/supabase-postgres-best-practices/SKILL.md +97 -97
  156. package/.agent/skills/swiftui-expert/SKILL.md +204 -201
  157. package/.agent/skills/system-design-pro/SKILL.md +345 -0
  158. package/.agent/skills/systematic-debugging/SKILL.md +153 -142
  159. package/.agent/skills/tailwind-patterns/SKILL.md +610 -566
  160. package/.agent/skills/tdd-workflow/SKILL.md +169 -161
  161. package/.agent/skills/test-result-analyzer/SKILL.md +313 -309
  162. package/.agent/skills/testing-patterns/SKILL.md +566 -579
  163. package/.agent/skills/trend-researcher/SKILL.md +243 -237
  164. package/.agent/skills/typescript-advanced/SKILL.md +336 -335
  165. package/.agent/skills/ui-ux-pro-max/SKILL.md +590 -562
  166. package/.agent/skills/ui-ux-researcher/SKILL.md +244 -244
  167. package/.agent/skills/vue-expert/SKILL.md +294 -275
  168. package/.agent/skills/vulnerability-scanner/SKILL.md +416 -404
  169. package/.agent/skills/web-accessibility-auditor/SKILL.md +219 -218
  170. package/.agent/skills/web-design-guidelines/SKILL.md +192 -186
  171. package/.agent/skills/webapp-testing/SKILL.md +167 -169
  172. package/.agent/skills/webgpu-performance/SKILL.md +56 -2
  173. package/.agent/skills/whimsy-injector/SKILL.md +346 -325
  174. package/.agent/skills/workflow-optimizer/SKILL.md +231 -229
  175. package/.agent/workflows/acf.md +141 -0
  176. package/.agent/workflows/api-tester.md +176 -151
  177. package/.agent/workflows/audit.md +150 -127
  178. package/.agent/workflows/brainstorm.md +134 -110
  179. package/.agent/workflows/changelog.md +140 -112
  180. package/.agent/workflows/create.md +168 -124
  181. package/.agent/workflows/debug.md +190 -165
  182. package/.agent/workflows/deploy.md +201 -180
  183. package/.agent/workflows/enhance.md +154 -128
  184. package/.agent/workflows/fix.md +136 -114
  185. package/.agent/workflows/generate.md +198 -183
  186. package/.agent/workflows/marathon.md +37 -11
  187. package/.agent/workflows/migrate.md +184 -160
  188. package/.agent/workflows/orchestrate.md +192 -168
  189. package/.agent/workflows/performance-benchmarker.md +135 -114
  190. package/.agent/workflows/plan.md +196 -173
  191. package/.agent/workflows/preview.md +103 -80
  192. package/.agent/workflows/refactor.md +192 -161
  193. package/.agent/workflows/review-ai.md +125 -101
  194. package/.agent/workflows/review.md +141 -116
  195. package/.agent/workflows/session.md +122 -94
  196. package/.agent/workflows/status.md +101 -79
  197. package/.agent/workflows/strengthen-skills.md +164 -138
  198. package/.agent/workflows/super-prompt.md +24 -0
  199. package/.agent/workflows/swarm.md +193 -179
  200. package/.agent/workflows/test.md +211 -189
  201. package/.agent/workflows/tribunal-backend.md +136 -105
  202. package/.agent/workflows/tribunal-database.md +122 -95
  203. package/.agent/workflows/tribunal-frontend.md +221 -96
  204. package/.agent/workflows/tribunal-full.md +129 -100
  205. package/.agent/workflows/tribunal-mobile.md +122 -95
  206. package/.agent/workflows/tribunal-performance.md +136 -110
  207. package/.agent/workflows/tribunal-speed.md +209 -183
  208. package/.agent/workflows/ui-ux-pro-max.md +145 -122
  209. package/README.md +107 -55
  210. package/bin/mcp-server.js +159 -0
  211. package/bin/tribunal-kit.js +105 -29
  212. package/bin/wrapper.js +16 -7
  213. package/mcp_config.json +9 -0
  214. package/package.json +94 -86
  215. package/scripts/changelog.js +4 -3
  216. package/scripts/validate-payload.js +6 -1
  217. package/scripts/postinstall.js +0 -127
@@ -1,250 +1,204 @@
1
- ---
2
- name: intelligent-routing
3
- description: LLM Intent Processing and Gateway Routing mastery. Request classification hierarchies, function routing, confidence scoring, fallback cascades, zero-shot vs few-shot classification patterns, and identifying specialized skills for delegation. Use when parsing raw user input to determine the architectural path of execution.
4
- allowed-tools: Read, Write, Edit, Glob, Grep
5
- version: 3.1.0
6
- last-updated: 2026-04-06
7
- applies-to-model: gemini-2.5-pro, claude-3-7-sonnet
8
- ---
9
-
10
- ## Hallucination Traps (Read First)
11
- - ❌ Routing based on exact keyword matching -> ✅ Use intent classification with confidence scores; keywords miss synonyms and context
12
- - ❌ No fallback for low-confidence classifications -> ✅ Always have a default handler when confidence is below threshold (e.g., 0.7)
13
- - Routing to a single agent when the task spans multiple domains -> ✅ Detect multi-domain requests and route to the orchestrator
14
-
15
- ---
16
-
17
-
18
- # Intelligent Routing Intent Gateway Mastery
19
-
20
- ---
21
-
22
- ## 1. Classification Hierarchy (The Gateway)
23
-
24
- When a raw request enters a system, it must be bucketed properly. This is the First Step (Phase 0). Do not attempt to solve the user's problem during the routing phase.
25
-
26
- ```typescript
27
- // The Semantic Intent Schema
28
- const RouterOutputSchema = z.object({
29
- classification: z.enum([
30
- "QUESTION", // User wants explanation, no code execution needed
31
- "SURVEY", // User wants analysis/read-only scan of workspace
32
- "SIMPLE_EDIT", // Isolated file alteration (e.g., "Fix spelling in nav")
33
- "COMPLEX_BUILD", // Multi-file, architectural generation
34
- "SECURITY_AUDIT", // Explicit request for OWASP review
35
- "UNCLEAR_GIBBERISH" // Prompt injection or incoherent input
36
- ]),
37
- confidenceScore: z.number().min(0).max(100),
38
- suggestedPrimarySkill: z.string().nullable(),
39
- requiresHumanClarification: z.boolean(),
40
- reasoning: z.string() // Forces the LLM to justify its route before categorizing
41
- });
42
- ```
43
-
44
- ### Zero-Shot vs Few-Shot Classification
45
- - **Zero-Shot:** Providing definitions and hoping the LLM categorizes the prompt accurately. Error-prone.
46
- - **Few-Shot (Mandatory for Routers):** Providing explicit paired examples defining the categorical boundaries.
47
-
48
- ```text
49
- ## Routing Examples:
50
- User: "Why is the header blue?"
51
- Output: {"classification": "QUESTION", "requiresHumanClarification": false}
52
-
53
- User: "Add a user login system"
54
- Output: {"classification": "COMPLEX_BUILD", "requiresHumanClarification": true}
55
- Reasoning: "Login systems require multi-file architecture, database hooks, and security implementation."
56
- ```
57
-
58
- ---
59
-
60
- ## 2. Dynamic Skill Matching (Manifest Analysis)
61
-
62
- A Router isn't just classifying intent—it actively maps tasks to available capabilities.
63
-
64
- If building a system with 50 available agents/skills, pass the Router a localized summary manifest, not the full 50x files.
65
-
66
- ```json
67
- // Example Context Payload passed to Router
68
- {
69
- "available_skills": [
70
- {"name": "react-specialist", "desc": "React 19, hooks, component architecture"},
71
- {"name": "python-pro", "desc": "FastAPI, async, data processing"},
72
- {"name": "vulnerability-scanner", "desc": "OWASP, injections, secret scanning"}
73
- ],
74
- "user_request": "How do I speed up this data pipeline script?"
75
- }
76
- ```
77
- *Router calculates:* `match: python-pro` AND `match: performance-profiling`.
78
-
79
- ---
80
-
81
- ## 3. Fallback Cascades & Ambiguity
82
-
83
- The AI will encounter prompts it does not understand. The Router is the *only* place where it is safe to halt and ask immediately.
84
-
85
- **The Socratic Yield Rule:**
86
- If the `confidenceScore` of a categorization is `< 85`, the router MUST yield back to the user with a clarifying question instead of guessing the intent.
87
-
88
- *User:* "Fix the thing."
89
- *Router Action (Incorrect):* Assume they mean standard linter execution and run scripts.
90
- *Router Action (Correct):* Halt. "Which file or feature are you referring to?"
91
-
92
- ---
93
-
94
- ## 4. Bounding the Exploder Pattern
95
-
96
- Certain requests sound simple but require massive execution matrices (The "Exploder" pattern).
97
- *User:* "Translate my entire app to French."
98
-
99
- The Router must recognize execution scales. If an execution requires touching >10 files, the Router must switch the system into `PLANNING_MODE` to generate an itinerary, rather than attempting an outright sequential execution.
100
-
101
- ---
102
-
103
- ## Intelligent Routing: Skill Manifest
104
- This file contains all available skills and workflows as a condensed index for the pre-router.
105
-
106
- |Skill Name|Description|
107
- |---|---|
108
- |`agent-organizer`|Senior agent organizer with expertise in assembling and coordinating multi-agent teams. Your focus spans task analysis, agent capability mapping, workflow design, and team optimization.|
109
- |`agentic-patterns`|AI agent design principles. Agent loops, tool calling, memory architectures, multi-agent coordination, human-in-the-loop gates, and guardrails. Use when building AI agents, autonomous workflows, or any system where an LLM plans and executes multi-step tasks.|
110
- |`api-patterns`|API design principles and decision-making. REST vs GraphQL vs tRPC selection, response formats, versioning, pagination.|
111
- |`app-builder`|Main application building orchestrator. Creates full-stack applications from natural language requests. Determines project type, selects tech stack, coordinates agents.|
112
- |`architecture`|Architectural decision-making framework. Requirements analysis, trade-off evaluation, ADR documentation. Use when making architecture decisions or analyzing system design.|
113
- |`bash-linux`|Bash/Linux terminal patterns. Critical commands, piping, error handling, scripting. Use when working on macOS or Linux systems.|
114
- |`behavioral-modes`|AI operational modes (brainstorm, implement, debug, review, teach, ship, orchestrate). Use to adapt behavior based on task type.|
115
- |`brainstorming`|Socratic questioning protocol + user communication. MANDATORY for complex requests, new features, or unclear requirements. Includes progress reporting and error handling.|
116
- |`clean-code`|Pragmatic coding standards - concise, direct, no over-engineering, no unnecessary comments|
117
- |`code-review-checklist`|Code review guidelines covering code quality, security, and best practices.|
118
- |`config-validator`|Self-validation skill for the .agent directory. Checks that all agents, skills, workflows, and scripts referenced across the system actually exist and are consistent. Use after modifying agent configuration files.|
119
- |`csharp-developer`|Senior C# developer with mastery of .NET 8+ and the Microsoft ecosystem. Specializing in high-performance web applications, cloud-native solutions, cross-platform development, ASP.NET Core, Blazor, and Entity Framework Core.|
120
- |`database-design`|Database design principles and decision-making. Schema design, indexing strategy, ORM selection, serverless databases.|
121
- |`deployment-procedures`|Production deployment principles and decision-making. Safe deployment workflows, rollback strategies, and verification. Teaches thinking, not scripts.|
122
- |`devops-engineer`|Senior DevOps engineer with expertise in building scalable, automated infrastructure and deployment pipelines. Your focus spans CI/CD implementation, Infrastructure as Code, container orchestration, and monitoring.|
123
- |`devops-incident-responder`|Senior DevOps incident responder with expertise in managing critical production incidents, performing rapid diagnostics, and implementing permanent fixes. Reduces MTTR and builds resilient systems.|
124
- |`documentation-templates`|Documentation templates and structure guidelines. README, API docs, code comments, and AI-friendly documentation.|
125
- |`dotnet-core-expert`|Senior .NET Core expert with expertise in .NET 10, C# 14, and modern minimal APIs. Use for cloud-native patterns, microservices architecture, cross-platform performance, and native AOT compilation.|
126
- |`edge-computing`|Edge function design principles. Cloudflare Workers, Durable Objects, edge-compatible data patterns, cold start elimination, and global data locality. Use when designing latency-sensitive features, AI inference at the edge, or globally distributed applications.|
127
- |`frontend-design`|Design thinking and decision-making for web UI. Use when designing components, layouts, color schemes, typography, or creating aesthetic interfaces. Teaches principles, not fixed values.|
128
- |`game-development`|Game development orchestrator. Routes to platform-specific skills based on project needs.|
129
- |`geo-fundamentals`|Generative Engine Optimization for AI search engines (ChatGPT, Claude, Perplexity).|
130
- |`i18n-localization`|Internationalization and localization patterns. Detecting hardcoded strings, managing translations, locale files, RTL support.|
131
- |`intelligent-routing`|Automatic agent selection and intelligent task routing. Analyzes user requests and automatically selects the best specialist agent(s) without requiring explicit user mentions.|
132
- |`lint-and-validate`|Linting and validation principles for code quality enforcement.|
133
- |`llm-engineering`|LLM engineering principles for production AI systems. RAG pipeline design, vector store selection, prompt engineering, evals, and LLMOps. Use when building AI features, chat interfaces, semantic search, or any system calling an LLM API.|
134
- |`local-first`|Local-first software principles. Offline-capable apps, CRDTs, sync engines (ElectricSQL, Replicache, Zero), conflict resolution, and the migration path from REST-first to local-first architecture. Use when building apps that need offline support, fast UI, or collaborative editing.|
135
- |`mcp-builder`|MCP (Model Context Protocol) server building principles. Tool design, resource patterns, best practices.|
136
- |`mobile-design`|Mobile-first and Spatial computing design thinking for iOS, Android, Foldables, and WebXR. Touch interaction, advanced haptics, on-device AI patterns, performance extremis. Teaches principles, not fixed values.|
137
- |`nextjs-react-expert`|Next.js App Router and React v19+ performance optimization from Vercel Engineering. Use when building React components, optimizing performance, implementing React Compiler patterns, eliminating waterfalls, reducing JS payload, or implementing Streaming/PPR optimizations.|
138
- |`nodejs-best-practices`|Node.js development principles and decision-making. Framework selection, async patterns, security, and architecture. Teaches thinking, not copying.|
139
- |`observability`|Production observability principles. OpenTelemetry traces, structured logs, metrics, SLOs/SLIs/error budgets, and AI observability. Use when setting up monitoring, debugging production issues, or designing observable distributed systems.|
140
- |`parallel-agents`|Multi-agent orchestration patterns. Use when multiple independent tasks can run with different domain expertise or when comprehensive analysis requires multiple perspectives.|
141
- |`performance-profiling`|Performance profiling principles. Measurement, analysis, and optimization techniques.|
142
- |`plan-writing`|Structured task planning with clear breakdowns, dependencies, and verification criteria. Use when implementing features, refactoring, or any multi-step work.|
143
- |`platform-engineer`|Senior platform engineer with deep expertise in building internal developer platforms, self-service infrastructure, and developer portals. Reduces cognitive load and accelerates software delivery.|
144
- |`powershell-windows`|PowerShell Windows patterns. Critical pitfalls, operator syntax, error handling.|
145
- |`python-patterns`|Python development principles and decision-making. Framework selection, async patterns, type hints, project structure. Teaches thinking, not copying.|
146
- |`python-pro`|Senior Python developer (3.11+) specializing in idiomatic, type-safe, and performant Python. Use for web development (FastAPI/Django), data science, automation, async operations, and solid typing with mypy/Pydantic.|
147
- |`react-specialist`|Senior React specialist (React 18+) focusing on advanced patterns, state management, performance optimization, and production architectures (Next.js/Remix).|
148
- |`realtime-patterns`|Real-time and collaborative application patterns. WebSockets, Server-Sent Events for AI streaming, CRDTs for conflict-free collaboration, presence, and sync engines. Use when building live collaboration, AI streaming UIs, live dashboards, or multiplayer features.|
149
- |`red-team-tactics`|Red team tactics principles based on MITRE ATT&CK. Attack phases, detection evasion, reporting.|
150
- |`rust-pro`|Master Rust 1.75+ with modern async patterns, advanced type system features, and production-ready systems programming. Expert in the latest Rust ecosystem including Tokio, axum, and cutting-edge crates. Use PROACTIVELY for Rust development, performance optimization, or systems programming.|
151
- |`seo-fundamentals`|SEO fundamentals, E-E-A-T, Core Web Vitals, and Google algorithm principles.|
152
- |`server-management`|Server management principles and decision-making. Process management, monitoring strategy, and scaling decisions. Teaches thinking, not commands.|
153
- |`sql-pro`|Senior SQL developer across major databases (PostgreSQL, MySQL, SQL Server, Oracle). Use for complex query design, performance optimization, indexing strategies, CTEs, window functions, and schema architecture.|
154
- |`systematic-debugging`|4-phase systematic debugging methodology with root cause analysis and evidence-based verification. Use when debugging complex issues.|
155
- |`tailwind-patterns`|Tailwind CSS v4+ principles for extreme frontend engineering. CSS-first configuration, scroll-driven animations, logical properties, advanced container style queries, and `@property` Houdini patterns.|
156
- |`tdd-workflow`|Test-Driven Development workflow principles. RED-GREEN-REFACTOR cycle.|
157
- |`test-result-analyzer`|Ingests test logs and identifies root causes across multiple failing test files. Provides actionable fix recommendations.|
158
- |`testing-patterns`|Testing patterns and principles. Unit, integration, mocking strategies.|
159
- |`trend-researcher`|Creative muse and design trend analyzer for modern web/mobile interfaces.|
160
- |`ui-ux-pro-max`|Plan and implement cutting-edge advanced UI/UX. Create distinctive, production-grade frontend interfaces with high design quality.|
161
- |`ui-ux-researcher`|Expert auditor for accessibility, cognitive load, and premium design heuristics.|
162
- |`vue-expert`|Vue 3 Composition API and modern Vue ecosystem expert. Use when building Vue applications, optimizing reactivity, component architecture, Nuxt 3 development, performance tuning, and State Management (Pinia).|
163
- |`vulnerability-scanner`|Advanced vulnerability analysis principles. OWASP 2025, Supply Chain Security, attack surface mapping, risk prioritization.|
164
- |`web-design-guidelines`|Review UI code for Next-Generation Web Interface Guidelines compliance. Use when asked to "review my UI", "check accessibility", "audit design", "review UX", or "check my site against best practices".|
165
- |`webapp-testing`|Web application testing principles. E2E, Playwright, deep audit strategies.|
166
- |`whimsy-injector`|Micro-delight generator for frontend interfaces. Suggests and implements subtle animations, playful transitions, and interaction polish across any frontend stack.|
167
- |`workflow-optimizer`|Analyzes agent tool-calling patterns and task execution efficiency to suggest process improvements.|
168
- |`api-architect`|API contract design agent. Builds robust REST/GraphQL/tRPC endpoints with RFC 9457 error formats, idempotency, pagination, and versioning. Use when designing new API routes or reviewing API contracts.|
169
- |`resilience-reviewer`|Fault tolerance reviewer. Audits for swallowed errors, naked Promises, missing retries/timeouts, absent circuit breakers, and missing React error boundaries. Use when reviewing backend or full-stack code for reliability.|
170
- |`schema-reviewer`|Input validation reviewer. Enforces strict Zod/Pydantic validation at all trust boundaries (API, env, query params). Catches `z.any()`, missing `.parse()`, and unvalidated external data. Use when reviewing data ingestion or API input handling.|
171
- |`gsap-core`|GSAP core animation API — gsap.to(), from(), fromTo(), easing, stagger, defaults, matchMedia(). Use for JavaScript animation with GSAP.|
172
- |`gsap-scrolltrigger`|GSAP ScrollTrigger plugin — scroll-driven animations, pinning, snapping. Use when building scroll-linked animation.|
173
- |`gsap-timeline`|GSAP timeline sequencing — Timeline API, position parameters, labels, nesting. Use for multi-step animation choreography.|
174
- |`gsap-react`|GSAP in React — useGSAP hook, ref-based targeting, cleanup, context scoping. Use when animating React components with GSAP.|
175
- |`gsap-plugins`|GSAP plugins — Flip, Draggable, MorphSVG, SplitText, MotionPath, DrawSVG. Use for advanced GSAP effects.|
176
- |`gsap-performance`|GSAP performance — will-change, GPU compositing, lazy MediaQuery, reduced motion. Use when optimizing GSAP animations.|
177
- |`gsap-utils`|GSAP utility methods — toArray, clamp, mapRange, interpolate, distribute. Use for GSAP helper functions.|
178
- |`gsap-frameworks`|GSAP framework integration — Vue, Svelte, Astro, Angular, Webflow. Use when integrating GSAP outside React.|
179
- |`error-resilience`|Error resilience patterns. Retry strategies, circuit breakers, graceful degradation, timeout policies.|
180
- |`data-validation-schemas`|Data validation with Zod, Pydantic, and JSON Schema. Use when building validation layers.|
181
- |`monorepo-management`|Monorepo tooling and workspace management. Turborepo, Nx, pnpm workspaces.|
182
- |`typescript-advanced`|Advanced TypeScript patterns. Generics, conditional types, template literals, discriminated unions.|
183
- |`game-developer`|Game development orchestrator. Routes to platform-specific skills (Unity/C#, Godot/GDScript, WebGL). Use when building games or game engines.|
184
- |`documentation-writer`|Technical documentation specialist. README files, API docs, code comments, architecture docs, and AI-friendly documentation. Use when writing or reviewing documentation.|
185
- |`test-engineer`|Test generation and strategy specialist. Creates tests using the Testing Trophy (unit → integration → E2E). Use when generating tests or designing test architecture.|
186
- |`qa-automation-engineer`|QA automation specialist. End-to-end testing pipelines, Playwright/Cypress configuration, CI test integration, visual regression. Use when building automated QA systems.|
187
- |`code-archaeologist`|Legacy code exploration specialist. Navigates unfamiliar codebases, identifies patterns, maps dependencies, documents tribal knowledge. Use when exploring or understanding legacy code.|
188
- |`project-planner`|Project planning specialist. 4-phase methodology (Analyze → Plan → Solution → Implement). Use when planning features, roadmaps, or multi-step implementations.|
189
- |`product-manager`|Product strategy specialist. Feature prioritization, roadmap planning, stakeholder alignment, competitive analysis. Use when making product decisions or prioritizing features.|
190
- |`product-owner`|User story and backlog management specialist. Writing acceptance criteria, sprint planning, backlog grooming. Use when defining user stories or managing backlogs.|
191
- |`seo-specialist`|SEO optimization specialist. Technical SEO, meta tags, structured data (JSON-LD), Core Web Vitals impact on search ranking, sitemap/robots configuration. Use when optimizing search visibility.|
192
- |`throughput-optimizer`|Throughput and latency optimization specialist. Load testing, connection pooling, queue management, batch processing, caching strategies. Use when optimizing system throughput or reducing latency.|
193
- |`vitals-reviewer`|Core Web Vitals specialist. LCP, CLS, INP measurement and optimization, Lighthouse auditing, real user monitoring. Use when auditing or improving web performance metrics.|
194
- |`penetration-tester`|Penetration testing and red team specialist. MITRE ATT&CK mapping, attack surface analysis, exploitation paths, vulnerability chaining. Use when pen testing applications or assessing attack surfaces.|
195
- |`db-latency-auditor`|Database performance specialist. Slow query analysis, EXPLAIN ANALYZE, index optimization, N+1 detection, connection pool tuning. Use when debugging database performance issues.|
196
- |`ai-code-reviewer`|AI/LLM integration code reviewer. Hallucinated model names, invented API parameters, prompt injection vulnerabilities, missing rate limits, streaming error handling. Use when reviewing code that calls LLM APIs.|
197
-
198
-
199
- ---
200
-
201
-
202
-
203
- AI coding assistants often fall into specific bad habits when dealing with this domain. These are strictly forbidden:
204
-
205
- 1. **Over-engineering:** Proposing complex abstractions or distributed systems when a simpler approach suffices.
206
- 2. **Hallucinated Libraries/Methods:** Using non-existent methods or packages. Always `// VERIFY` or check `package.json` / `requirements.txt`.
207
- 3. **Skipping Edge Cases:** Writing the "happy path" and ignoring error handling, timeouts, or data validation.
208
- 4. **Context Amnesia:** Forgetting the user's constraints and offering generic advice instead of tailored solutions.
209
- 5. **Silent Degradation:** Catching and suppressing errors without logging or re-raising.
210
-
211
- ---
212
-
213
-
214
-
215
- **Slash command: `/review` or `/tribunal-full`**
216
- **Active reviewers: `logic-reviewer` · `security-auditor`**
217
-
218
- ### ❌ Forbidden AI Tropes
219
-
220
- 1. **Blind Assumptions:** Never make an assumption without documenting it clearly with `// VERIFY: [reason]`.
221
- 2. **Silent Degradation:** Catching and suppressing errors without logging or handling.
222
- 3. **Context Amnesia:** Forgetting the user's constraints and offering generic advice instead of tailored solutions.
223
-
224
-
225
-
226
- Review these questions before confirming output:
227
- ```
228
- ✅ Did I rely ONLY on real, verified tools and methods?
229
- ✅ Is this solution appropriately scoped to the user's constraints?
230
- ✅ Did I handle potential failure modes and edge cases?
231
- ✅ Have I avoided generic boilerplate that doesn't add value?
232
- ```
233
-
234
- ### 🛑 Verification-Before-Completion (VBC) Protocol
235
-
236
- **CRITICAL:** You must follow a strict "evidence-based closeout" state machine.
237
- - ❌ **Forbidden:** Declaring a task complete because the output "looks correct."
238
- - ✅ **Required:** You are explicitly forbidden from finalizing any task without providing **concrete evidence** (terminal output, passing tests, compile success, or equivalent proof) that your output works as intended.
239
-
240
-
241
- ## Pre-Flight Checklist
242
- - [ ] Have I reviewed the user's specific constraints and requests?
243
- - [ ] Have I checked the environment for relevant existing implementations?
244
-
245
- ## VBC Protocol (Verification-Before-Completion)
246
- You MUST verify existing code signatures and variables before attempting to modify or call them. No hallucination is permitted.
1
+ ---
2
+ name: intelligent-routing
3
+ description: LLM Intent Processing and Gateway Routing mastery. Request classification hierarchies, function routing, confidence scoring, fallback cascades, zero-shot vs few-shot classification patterns, and identifying specialized skills for delegation. Use when parsing raw user input to determine the architectural path of execution.
4
+ allowed-tools: Read, Write, Edit, Glob, Grep
5
+ version: 4.0.0
6
+ last-updated: 2026-06-21
7
+ applies-to-model: gemini-2.5-pro, claude-3-7-sonnet
8
+ routing:
9
+ domain: meta
10
+ tier: basic
11
+ ---
12
+
13
+ ## Hallucination Traps (Read First)
14
+
15
+ - ❌ Routing based on exact keyword matching -> ✅ Use intent classification with confidence scores; keywords miss synonyms and context
16
+ - ❌ No fallback for low-confidence classifications -> ✅ Always have a default handler when confidence is below threshold (e.g., 0.7)
17
+ - ❌ Routing to a single agent when the task spans multiple domains -> ✅ Detect multi-domain requests and route to the orchestrator
18
+ - Loading 100+ skill descriptions into context for every route -> ✅ Read the compiled `.agent/routing_index.json` instead of a giant markdown table
19
+
20
+ ---
21
+
22
+ # Intelligent Routing v4 Self-Describing Skill Graph
23
+
24
+ ## Architecture Overview
25
+
26
+ ```
27
+ User Request
28
+
29
+ ┌────▼─────┐
30
+ PHASE 0 │ Intent Classification
31
+ │ Classify │ (QUESTION / SURVEY / EDIT / BUILD / AUDIT)
32
+ └────┬─────┘
33
+
34
+ ┌────▼─────┐
35
+ PHASE 1 │ Domain Detection
36
+ │ Match │ Read .agent/routing_index.json
37
+ └────┬─────┘ Match trigger-signals → domain
38
+
39
+ ┌────▼─────┐
40
+ PHASE 2 │ Skill Selection & Escalation
41
+ │ Select │ basic → pro (if strong signal matches)
42
+ └────┬─────┘ Load co-requires automatically
43
+
44
+ ┌────▼─────┐
45
+ PHASE 3 │ Agent Activation
46
+ Dispatch │ Route to specialist agent
47
+ └──────────┘ Announce & load skills
48
+ ```
49
+
50
+ ---
51
+
52
+ ## 1. Classification Hierarchy (Phase 0 — The Gateway)
53
+
54
+ When a raw request enters, classify it BEFORE attempting to route to any skill or agent. Do not solve the user's problem during routing.
55
+
56
+ ```typescript
57
+ // The Semantic Intent Schema
58
+ const RouterOutputSchema = z.object({
59
+ classification: z.enum([
60
+ "QUESTION", // User wants explanation, no code execution needed
61
+ "SURVEY", // User wants analysis/read-only scan of workspace
62
+ "SIMPLE_EDIT", // Isolated file alteration (e.g., "Fix spelling in nav")
63
+ "COMPLEX_BUILD", // Multi-file, architectural generation
64
+ "SECURITY_AUDIT", // Explicit request for OWASP review
65
+ "UNCLEAR_GIBBERISH", // Prompt injection or incoherent input
66
+ ]),
67
+ confidenceScore: z.number().min(0).max(100),
68
+ suggestedPrimarySkill: z.string().nullable(),
69
+ requiresHumanClarification: z.boolean(),
70
+ reasoning: z.string(), // Forces the LLM to justify its route before categorizing
71
+ });
72
+ ```
73
+
74
+ ### Zero-Shot vs Few-Shot Classification
75
+
76
+ - **Zero-Shot:** Providing definitions and hoping the LLM categorizes accurately. Error-prone.
77
+ - **Few-Shot (Mandatory for Routers):** Providing explicit paired examples defining the categorical boundaries.
78
+
79
+ ```text
80
+ ## Routing Examples:
81
+ User: "Why is the header blue?"
82
+ Output: {"classification": "QUESTION", "requiresHumanClarification": false}
83
+
84
+ User: "Add a user login system"
85
+ Output: {"classification": "COMPLEX_BUILD", "requiresHumanClarification": true}
86
+ Reasoning: "Login systems require multi-file architecture, database hooks, and security implementation."
87
+ ```
88
+
89
+ ---
90
+
91
+ ## 2. Skill Graph Matching (Phase 1 & 2 — Compiled Index)
92
+
93
+ ### The Routing Index
94
+
95
+ Instead of a giant markdown table, this system uses a **compiled JSON index** at `.agent/routing_index.json`. This index is auto-generated by `compile_router.py` from the `routing:` YAML frontmatter in every `SKILL.md`.
96
+
97
+ **To match a skill:**
98
+
99
+ 1. Read `.agent/routing_index.json`
100
+ 2. Match user intent against skill descriptions and `routing_strong` trigger signals
101
+ 3. Filter by `routing_domain` to narrow candidates
102
+ 4. Apply escalation rules (see below)
103
+
104
+ ### Skill Frontmatter Schema
105
+
106
+ Every skill declares its routing metadata in its YAML frontmatter:
107
+
108
+ ```yaml
109
+ routing:
110
+ domain: devops | frontend | backend | architecture | data | security | testing | design | meta | general
111
+ tier: basic | pro
112
+ supersedes: <skill-name> # "I replace this basic skill for advanced use"
113
+ co-requires: [<skill>, ...] # "Load these alongside me"
114
+ conflicts-with: [<skill>, ...] # "Don't load both"
115
+ trigger-signals:
116
+ strong: [keyword1, keyword2] # High-confidence activation triggers
117
+ weak: [keyword3, keyword4] # Low-confidence, need additional context
118
+ confidence-boost: <number> # How much to boost score when strong signal matches
119
+ ```
120
+
121
+ ### Escalation Rules (basic pro)
247
122
 
123
+ When a user's request contains **strong trigger signals** that match a `tier: pro` skill:
124
+
125
+ ```
126
+ 1. Check if any tier:pro skill's strong signals match the request
127
+ 2. If YES and the pro skill has `supersedes: <basic-skill>`:
128
+ → Load the pro skill INSTEAD of the basic one
129
+ → Example: "OIDC GitHub Actions" → git-pro (supersedes github-operations)
130
+ 3. If YES and the pro skill has `co-requires`:
131
+ → Also load the co-required skills
132
+ → Example: cicd-pro co-requires [containerization-pro, cloud-architect]
133
+ ```
134
+
135
+ ### Conflict Resolution
136
+
137
+ When multiple skills match:
138
+
139
+ ```
140
+ Priority order:
141
+ 1. Exact strong signal match > weak signal match
142
+ 2. tier:pro > tier:basic (when both match)
143
+ 3. Specific domain > general domain
144
+ 4. If a pro skill `supersedes` a basic skill, drop the basic skill
145
+ 5. If two skills have `conflicts-with` each other, pick the one with higher signal match count
146
+ ```
147
+
148
+ ---
149
+
150
+ ## 3. Fallback Cascades & Ambiguity
151
+
152
+ The AI will encounter prompts it does not understand. The Router is the _only_ place where it is safe to halt and ask immediately.
153
+
154
+ **The Socratic Yield Rule:**
155
+ If the `confidenceScore` of a categorization is `< 85`, the router MUST yield back to the user with a clarifying question instead of guessing the intent.
156
+
157
+ _User:_ "Fix the thing."
158
+ _Router Action (Incorrect):_ Assume they mean standard linter execution and run scripts.
159
+ _Router Action (Correct):_ Halt. "Which file or feature are you referring to?"
160
+
161
+ ---
162
+
163
+ ## 4. Bounding the Exploder Pattern
164
+
165
+ Certain requests sound simple but require massive execution matrices (The "Exploder" pattern).
166
+ _User:_ "Translate my entire app to French."
167
+
168
+ The Router must recognize execution scales. If an execution requires touching >10 files, the Router must switch the system into `PLANNING_MODE` to generate an itinerary, rather than attempting an outright sequential execution.
169
+
170
+ ---
171
+
172
+ ## 5. Domain Overlap Disambiguation
173
+
174
+ When keywords belong to multiple domains, use these explicit rules:
175
+
176
+ | Signal Combination | Route To | NOT |
177
+ | ------------------------------------ | -------------------------------------------------- | -------------------- |
178
+ | Docker + AWS/ECR/Terraform | `cloud-engineer` | `devops-engineer` |
179
+ | Docker alone (local dev) | `devops-engineer` | `cloud-engineer` |
180
+ | Git + OIDC/monorepo/semantic-release | `git-pro` (via system-architect or cloud-engineer) | `github-operations` |
181
+ | Git + basic branching/commits | `github-operations` | `git-pro` |
182
+ | CI/CD + AWS ECS + deploy | `cloud-engineer` (loads cicd-pro) | `devops-engineer` |
183
+ | CI/CD + general GitHub Actions | `devops-engineer` | `cloud-engineer` |
184
+ | System design + scale + capacity | `system-architect` | `backend-specialist` |
185
+ | Architecture + code patterns | `backend-specialist` | `system-architect` |
186
+
187
+ ---
188
+
189
+ ## 6. Regenerating the Index
190
+
191
+ When new skills are added or existing frontmatter is modified, regenerate the index:
192
+
193
+ ```bash
194
+ python .agent/scripts/compile_router.py
195
+ ```
196
+
197
+ This is idempotent — running it multiple times produces the same output. The index should be regenerated after:
198
+
199
+ - Adding a new skill
200
+ - Modifying a skill's `routing:` frontmatter
201
+ - Removing a skill
248
202
 
249
203
  ---
250
204
 
@@ -274,6 +228,7 @@ AI coding assistants often fall into specific bad habits when dealing with this
274
228
  ### ✅ Pre-Flight Self-Audit
275
229
 
276
230
  Review these questions before confirming output:
231
+
277
232
  ```
278
233
  ✅ Did I rely ONLY on real, verified tools and methods?
279
234
  ✅ Is this solution appropriately scoped to the user's constraints?
@@ -284,5 +239,6 @@ Review these questions before confirming output:
284
239
  ### 🛑 Verification-Before-Completion (VBC) Protocol
285
240
 
286
241
  **CRITICAL:** You must follow a strict "evidence-based closeout" state machine.
242
+
287
243
  - ❌ **Forbidden:** Declaring a task complete because the output "looks correct."
288
244
  - ✅ **Required:** You are explicitly forbidden from finalizing any task without providing **concrete evidence** (terminal output, passing tests, compile success, or equivalent proof) that your output works as intended.
@@ -1,56 +1,62 @@
1
- ---
2
- name: Knowledge Graph Analyzer
3
- description: Understands the architecture, risk blast radius, and dependencies of the codebase without token bloat. Now includes Context Snapshots for 27x token reduction.
4
- version: 3.0.0
5
- ---
6
-
7
- # /graph — Knowledge Graph Skill v3.0
8
-
9
- Use this skill when the user types `/graph` or when you need to deeply understand the architecture of an unfamiliar codebase without suffering from context-window bloat.
10
-
11
- ## The Token Reduction Protocol (Option C)
12
-
13
- **DO NOT READ RAW SOURCE FILES IF A SNAPSHOT EXISTS.**
14
- Reading raw project files wastes up to 50,000 tokens per edit. You must use the pre-computed Context Snapshots instead. These snapshots contain the file's content, resolved imports, dependent files, and risk scores all in one JSON blob.
15
-
16
- ## Pre-Flight Checklist
17
- - [ ] Have I run the Macro Mapper to generate the latest Context Snapshots?
18
- - [ ] Am I reading from `.agent/history/snapshots/` instead of directly grepping the project?
19
- - [ ] Have I respected the `blastRadius` before modifying a file?
20
-
21
- ## Execution Protocol
22
-
23
- 1. **Step 1: The Macro Map (Blast Radius & Snapshot Engine)**
24
- Execute the graph builder to map module boundaries and compute downstream risk scores. This also automatically generates a Context Snapshot for every file.
25
- ```bash
26
- node .agent/scripts/graph_builder.js
27
- ```
28
-
29
- 2. **Step 2: Read Context Snapshots (MANDATORY)**
30
- Instead of using `cat` or `grep` to read a file, read its snapshot. Snapshots are stored in `.agent/history/snapshots/` with slashes replaced by `__`.
31
- *Example:* To edit `src/middleware/auth.js`, read `.agent/history/snapshots/src__middleware__auth.js.json`.
32
-
33
- This gives you:
34
- - The full source code of the target file.
35
- - The exported symbols of every file it imports.
36
- - The list of files that depend on it.
37
- - Its exact `riskScore` and `blastRadius`.
38
-
39
- 3. **Step 3: Interactive Visualization (For Humans)**
40
- The user can view a sleek, zero-dependency visualizer of the codebase. You can prompt them to run:
41
- ```bash
42
- npx tribunal-kit graph
43
- ```
44
-
45
- 4. **Step 4: The Micro Zoom (Legacy Street View)**
46
- If a snapshot is unavailable or too large, you can fall back to the zoomer to get its structural skeleton:
47
- ```bash
48
- node .agent/scripts/graph_zoom.js --focus <path_to_file>
49
- ```
50
-
51
- ## VBC Protocol (Verification-Before-Completion)
52
- You are explicitly forbidden from guessing or "hallucinating" what functions, props, or variables exist inside a file. You MUST read the Context Snapshot (or use `graph_zoom.js`) to verify a component's exact signature before you attempt to call it, mock it, or rewrite it. Always respect the Blast Radius Risk Score before deleting or mutating files.
1
+ ---
2
+ name: Knowledge Graph Analyzer
3
+ description: Understands the architecture, risk blast radius, and dependencies of the codebase without token bloat. Now includes Context Snapshots for 27x token reduction.
4
+ version: 3.0.0
5
+ routing:
6
+ domain: general
7
+ tier: basic
8
+ ---
9
+
10
+ # /graph — Knowledge Graph Skill v3.0
11
+
12
+ Use this skill when the user types `/graph` or when you need to deeply understand the architecture of an unfamiliar codebase without suffering from context-window bloat.
13
+
14
+ ## The Token Reduction Protocol (Option C)
15
+
16
+ **DO NOT READ RAW SOURCE FILES IF A SNAPSHOT EXISTS.**
17
+ Reading raw project files wastes up to 50,000 tokens per edit. You must use the pre-computed Context Snapshots instead. These snapshots contain the file's content, resolved imports, dependent files, and risk scores all in one JSON blob.
18
+
19
+ ## Pre-Flight Checklist
20
+
21
+ - [ ] Have I run the Macro Mapper to generate the latest Context Snapshots?
22
+ - [ ] Am I reading from `.agent/history/snapshots/` instead of directly grepping the project?
23
+ - [ ] Have I respected the `blastRadius` before modifying a file?
24
+
25
+ ## Execution Protocol
26
+
27
+ 1. **Step 1: The Macro Map (Blast Radius & Snapshot Engine)**
28
+ Execute the graph builder to map module boundaries and compute downstream risk scores. This also automatically generates a Context Snapshot for every file.
53
29
 
30
+ ```bash
31
+ node .agent/scripts/graph_builder.js
32
+ ```
33
+
34
+ 2. **Step 2: Read Context Snapshots (MANDATORY)**
35
+ Instead of using `cat` or `grep` to read a file, read its snapshot. Snapshots are stored in `.agent/history/snapshots/` with slashes replaced by `__`.
36
+ _Example:_ To edit `src/middleware/auth.js`, read `.agent/history/snapshots/src__middleware__auth.js.json`.
37
+
38
+ This gives you:
39
+ - The full source code of the target file.
40
+ - The exported symbols of every file it imports.
41
+ - The list of files that depend on it.
42
+ - Its exact `riskScore` and `blastRadius`.
43
+
44
+ 3. **Step 3: Interactive Visualization (For Humans)**
45
+ The user can view a sleek, zero-dependency visualizer of the codebase. You can prompt them to run:
46
+
47
+ ```bash
48
+ npx tribunal-kit graph
49
+ ```
50
+
51
+ 4. **Step 4: The Micro Zoom (Legacy Street View)**
52
+ If a snapshot is unavailable or too large, you can fall back to the zoomer to get its structural skeleton:
53
+ ```bash
54
+ node .agent/scripts/graph_zoom.js --focus <path_to_file>
55
+ ```
56
+
57
+ ## VBC Protocol (Verification-Before-Completion)
58
+
59
+ You are explicitly forbidden from guessing or "hallucinating" what functions, props, or variables exist inside a file. You MUST read the Context Snapshot (or use `graph_zoom.js`) to verify a component's exact signature before you attempt to call it, mock it, or rewrite it. Always respect the Blast Radius Risk Score before deleting or mutating files.
54
60
 
55
61
  ---
56
62
 
@@ -80,6 +86,7 @@ AI coding assistants often fall into specific bad habits when dealing with this
80
86
  ### ✅ Pre-Flight Self-Audit
81
87
 
82
88
  Review these questions before confirming output:
89
+
83
90
  ```
84
91
  ✅ Did I rely ONLY on real, verified tools and methods?
85
92
  ✅ Is this solution appropriately scoped to the user's constraints?
@@ -90,5 +97,6 @@ Review these questions before confirming output:
90
97
  ### 🛑 Verification-Before-Completion (VBC) Protocol
91
98
 
92
99
  **CRITICAL:** You must follow a strict "evidence-based closeout" state machine.
100
+
93
101
  - ❌ **Forbidden:** Declaring a task complete because the output "looks correct."
94
102
  - ✅ **Required:** You are explicitly forbidden from finalizing any task without providing **concrete evidence** (terminal output, passing tests, compile success, or equivalent proof) that your output works as intended.