tribunal-kit 4.5.1 → 4.6.1

This diff represents the content of publicly available package versions that have been released to one of the supported registries. The information contained in this diff is provided for informational purposes only and reflects changes between package versions as they appear in their respective public registries.
Files changed (214) hide show
  1. package/.agent/.shared/ui-ux-pro-max/README.md +4 -4
  2. package/.agent/ARCHITECTURE.md +282 -277
  3. package/.agent/agents/accessibility-reviewer.md +187 -187
  4. package/.agent/agents/ai-code-reviewer.md +199 -199
  5. package/.agent/agents/api-architect.md +71 -66
  6. package/.agent/agents/backend-specialist.md +219 -215
  7. package/.agent/agents/cloud-engineer.md +98 -0
  8. package/.agent/agents/code-archaeologist.md +168 -161
  9. package/.agent/agents/database-architect.md +184 -184
  10. package/.agent/agents/db-latency-auditor.md +213 -216
  11. package/.agent/agents/debugger.md +198 -191
  12. package/.agent/agents/dependency-reviewer.md +106 -103
  13. package/.agent/agents/devops-engineer.md +218 -218
  14. package/.agent/agents/documentation-writer.md +209 -201
  15. package/.agent/agents/explorer-agent.md +167 -160
  16. package/.agent/agents/frontend-reviewer.md +162 -160
  17. package/.agent/agents/frontend-specialist.md +257 -248
  18. package/.agent/agents/game-developer.md +48 -48
  19. package/.agent/agents/logic-reviewer.md +118 -116
  20. package/.agent/agents/mobile-developer.md +197 -200
  21. package/.agent/agents/mobile-reviewer.md +159 -162
  22. package/.agent/agents/orchestrator.md +187 -181
  23. package/.agent/agents/penetration-tester.md +160 -157
  24. package/.agent/agents/performance-optimizer.md +183 -183
  25. package/.agent/agents/performance-reviewer.md +178 -178
  26. package/.agent/agents/precedence-reviewer.md +251 -250
  27. package/.agent/agents/product-manager.md +149 -142
  28. package/.agent/agents/product-owner.md +81 -80
  29. package/.agent/agents/project-planner.md +152 -142
  30. package/.agent/agents/qa-automation-engineer.md +216 -225
  31. package/.agent/agents/resilience-reviewer.md +88 -88
  32. package/.agent/agents/schema-reviewer.md +67 -67
  33. package/.agent/agents/security-auditor.md +180 -174
  34. package/.agent/agents/seo-specialist.md +188 -193
  35. package/.agent/agents/sql-reviewer.md +159 -161
  36. package/.agent/agents/supervisor-agent.md +173 -184
  37. package/.agent/agents/swarm-worker-contracts.md +170 -166
  38. package/.agent/agents/swarm-worker-registry.md +92 -92
  39. package/.agent/agents/system-architect.md +85 -0
  40. package/.agent/agents/test-coverage-reviewer.md +158 -160
  41. package/.agent/agents/test-engineer.md +118 -118
  42. package/.agent/agents/throughput-optimizer.md +291 -299
  43. package/.agent/agents/type-safety-reviewer.md +182 -175
  44. package/.agent/agents/ui-ux-auditor.md +300 -292
  45. package/.agent/agents/vitals-reviewer.md +223 -223
  46. package/.agent/mcp_config.json +37 -40
  47. package/.agent/patterns/generator.md +11 -9
  48. package/.agent/patterns/inversion.md +14 -12
  49. package/.agent/patterns/pipeline.md +11 -9
  50. package/.agent/patterns/reviewer.md +15 -13
  51. package/.agent/patterns/tool-wrapper.md +11 -9
  52. package/.agent/routing_index.json +714 -0
  53. package/.agent/rules/GEMINI.md +359 -352
  54. package/.agent/scripts/compile_router.py +112 -0
  55. package/.agent/scripts/migrate_skills_frontmatter.py +64 -0
  56. package/.agent/scripts/strengthen_skills.js +1 -1
  57. package/.agent/skills/advanced-rag-pipelines/SKILL.md +56 -0
  58. package/.agent/skills/agent-organizer/SKILL.md +156 -150
  59. package/.agent/skills/agentic-patterns/SKILL.md +313 -315
  60. package/.agent/skills/ai-prompt-injection-defense/SKILL.md +190 -184
  61. package/.agent/skills/api-patterns/SKILL.md +253 -247
  62. package/.agent/skills/api-security-auditor/SKILL.md +195 -193
  63. package/.agent/skills/app-builder/SKILL.md +573 -572
  64. package/.agent/skills/app-builder/templates/SKILL.md +108 -115
  65. package/.agent/skills/app-builder/templates/astro-static/TEMPLATE.md +76 -76
  66. package/.agent/skills/app-builder/templates/chrome-extension/TEMPLATE.md +92 -92
  67. package/.agent/skills/app-builder/templates/cli-tool/TEMPLATE.md +88 -88
  68. package/.agent/skills/app-builder/templates/electron-desktop/TEMPLATE.md +88 -88
  69. package/.agent/skills/app-builder/templates/express-api/TEMPLATE.md +83 -83
  70. package/.agent/skills/app-builder/templates/flutter-app/TEMPLATE.md +90 -90
  71. package/.agent/skills/app-builder/templates/monorepo-turborepo/TEMPLATE.md +90 -90
  72. package/.agent/skills/app-builder/templates/nextjs-fullstack/TEMPLATE.md +126 -122
  73. package/.agent/skills/app-builder/templates/nextjs-saas/TEMPLATE.md +127 -122
  74. package/.agent/skills/app-builder/templates/nextjs-static/TEMPLATE.md +172 -169
  75. package/.agent/skills/app-builder/templates/nuxt-app/TEMPLATE.md +139 -134
  76. package/.agent/skills/app-builder/templates/python-fastapi/TEMPLATE.md +83 -83
  77. package/.agent/skills/app-builder/templates/react-native-app/TEMPLATE.md +122 -119
  78. package/.agent/skills/appflow-wireframe/SKILL.md +146 -145
  79. package/.agent/skills/architecture/SKILL.md +226 -219
  80. package/.agent/skills/authentication-best-practices/SKILL.md +197 -189
  81. package/.agent/skills/backend-security-expert/SKILL.md +16 -2
  82. package/.agent/skills/bash-linux/SKILL.md +179 -179
  83. package/.agent/skills/behavioral-modes/SKILL.md +239 -223
  84. package/.agent/skills/brainstorming/SKILL.md +498 -486
  85. package/.agent/skills/browser-native-ai/SKILL.md +57 -4
  86. package/.agent/skills/building-native-ui/SKILL.md +202 -202
  87. package/.agent/skills/cicd-pro/SKILL.md +442 -0
  88. package/.agent/skills/clean-code/SKILL.md +400 -381
  89. package/.agent/skills/cloud-architect/SKILL.md +439 -0
  90. package/.agent/skills/code-review-checklist/SKILL.md +203 -194
  91. package/.agent/skills/config-validator/SKILL.md +165 -165
  92. package/.agent/skills/containerization-pro/SKILL.md +452 -0
  93. package/.agent/skills/csharp-developer/SKILL.md +518 -518
  94. package/.agent/skills/data-validation-schemas/SKILL.md +333 -328
  95. package/.agent/skills/database-design/SKILL.md +247 -240
  96. package/.agent/skills/deployment-procedures/SKILL.md +172 -169
  97. package/.agent/skills/devops-engineer/SKILL.md +345 -345
  98. package/.agent/skills/devops-incident-responder/SKILL.md +143 -137
  99. package/.agent/skills/documentation-templates/SKILL.md +291 -279
  100. package/.agent/skills/edge-computing/SKILL.md +183 -181
  101. package/.agent/skills/emil-design-eng/SKILL.md +147 -0
  102. package/.agent/skills/error-resilience/SKILL.md +411 -428
  103. package/.agent/skills/extract-design-system/SKILL.md +160 -158
  104. package/.agent/skills/framer-motion-expert/SKILL.md +253 -244
  105. package/.agent/skills/frontend-design/SKILL.md +208 -201
  106. package/.agent/skills/frontend-security-expert/SKILL.md +16 -3
  107. package/.agent/skills/game-design-expert/SKILL.md +132 -129
  108. package/.agent/skills/game-engineering-expert/SKILL.md +148 -146
  109. package/.agent/skills/generative-ui-expert/SKILL.md +57 -1
  110. package/.agent/skills/geo-fundamentals/SKILL.md +148 -147
  111. package/.agent/skills/git-pro/SKILL.md +435 -0
  112. package/.agent/skills/github-operations/SKILL.md +335 -329
  113. package/.agent/skills/gsap-core/SKILL.md +319 -308
  114. package/.agent/skills/gsap-frameworks/SKILL.md +213 -207
  115. package/.agent/skills/gsap-performance/SKILL.md +139 -133
  116. package/.agent/skills/gsap-plugins/SKILL.md +486 -480
  117. package/.agent/skills/gsap-react/SKILL.md +202 -189
  118. package/.agent/skills/gsap-scrolltrigger/SKILL.md +357 -350
  119. package/.agent/skills/gsap-timeline/SKILL.md +165 -161
  120. package/.agent/skills/gsap-utils/SKILL.md +344 -338
  121. package/.agent/skills/harness-protocol/SKILL.md +48 -0
  122. package/.agent/skills/i18n-localization/SKILL.md +174 -163
  123. package/.agent/skills/intelligent-routing/SKILL.md +202 -246
  124. package/.agent/skills/knowledge-graph/SKILL.md +60 -52
  125. package/.agent/skills/lint-and-validate/SKILL.md +261 -261
  126. package/.agent/skills/llm-engineering/SKILL.md +400 -394
  127. package/.agent/skills/local-first/SKILL.md +178 -178
  128. package/.agent/skills/mcp-builder/SKILL.md +143 -142
  129. package/.agent/skills/mobile-design/SKILL.md +272 -263
  130. package/.agent/skills/monorepo-management/SKILL.md +335 -334
  131. package/.agent/skills/motion-engineering/SKILL.md +266 -234
  132. package/.agent/skills/nextjs-react-expert/SKILL.md +236 -234
  133. package/.agent/skills/nodejs-best-practices/SKILL.md +547 -548
  134. package/.agent/skills/observability/SKILL.md +343 -343
  135. package/.agent/skills/parallel-agents/SKILL.md +143 -146
  136. package/.agent/skills/performance-profiling/SKILL.md +259 -267
  137. package/.agent/skills/plan-writing/SKILL.md +150 -142
  138. package/.agent/skills/platform-engineer/SKILL.md +148 -147
  139. package/.agent/skills/playwright-best-practices/SKILL.md +188 -187
  140. package/.agent/skills/powershell-windows/SKILL.md +162 -162
  141. package/.agent/skills/project-idioms/SKILL.md +137 -137
  142. package/.agent/skills/python-patterns/SKILL.md +260 -259
  143. package/.agent/skills/python-pro/SKILL.md +324 -323
  144. package/.agent/skills/react-specialist/SKILL.md +305 -277
  145. package/.agent/skills/readme-builder/SKILL.md +310 -300
  146. package/.agent/skills/realtime-patterns/SKILL.md +323 -319
  147. package/.agent/skills/red-team-tactics/SKILL.md +231 -218
  148. package/.agent/skills/review-animations/SKILL.md +72 -0
  149. package/.agent/skills/review-animations/STANDARDS.md +73 -0
  150. package/.agent/skills/rust-pro/SKILL.md +671 -673
  151. package/.agent/skills/seo-fundamentals/SKILL.md +179 -179
  152. package/.agent/skills/server-management/SKILL.md +218 -214
  153. package/.agent/skills/shadcn-ui-expert/SKILL.md +231 -231
  154. package/.agent/skills/skill-creator/SKILL.md +87 -86
  155. package/.agent/skills/sql-pro/SKILL.md +629 -629
  156. package/.agent/skills/supabase-postgres-best-practices/SKILL.md +97 -97
  157. package/.agent/skills/swiftui-expert/SKILL.md +204 -201
  158. package/.agent/skills/system-design-pro/SKILL.md +345 -0
  159. package/.agent/skills/systematic-debugging/SKILL.md +153 -142
  160. package/.agent/skills/tailwind-patterns/SKILL.md +610 -566
  161. package/.agent/skills/tdd-workflow/SKILL.md +169 -161
  162. package/.agent/skills/test-result-analyzer/SKILL.md +313 -309
  163. package/.agent/skills/testing-patterns/SKILL.md +566 -579
  164. package/.agent/skills/trend-researcher/SKILL.md +243 -237
  165. package/.agent/skills/typescript-advanced/SKILL.md +336 -335
  166. package/.agent/skills/ui-ux-pro-max/SKILL.md +590 -562
  167. package/.agent/skills/ui-ux-researcher/SKILL.md +244 -244
  168. package/.agent/skills/vue-expert/SKILL.md +294 -275
  169. package/.agent/skills/vulnerability-scanner/SKILL.md +416 -404
  170. package/.agent/skills/web-accessibility-auditor/SKILL.md +219 -218
  171. package/.agent/skills/web-design-guidelines/SKILL.md +192 -186
  172. package/.agent/skills/webapp-testing/SKILL.md +167 -169
  173. package/.agent/skills/webgpu-performance/SKILL.md +56 -2
  174. package/.agent/skills/whimsy-injector/SKILL.md +346 -325
  175. package/.agent/skills/workflow-optimizer/SKILL.md +231 -229
  176. package/.agent/workflows/acf.md +141 -0
  177. package/.agent/workflows/api-tester.md +176 -151
  178. package/.agent/workflows/audit.md +150 -127
  179. package/.agent/workflows/brainstorm.md +134 -110
  180. package/.agent/workflows/changelog.md +140 -112
  181. package/.agent/workflows/create.md +168 -124
  182. package/.agent/workflows/debug.md +190 -165
  183. package/.agent/workflows/deploy.md +201 -180
  184. package/.agent/workflows/enhance.md +154 -128
  185. package/.agent/workflows/fix.md +136 -114
  186. package/.agent/workflows/generate.md +198 -183
  187. package/.agent/workflows/marathon.md +37 -11
  188. package/.agent/workflows/migrate.md +184 -160
  189. package/.agent/workflows/orchestrate.md +192 -168
  190. package/.agent/workflows/performance-benchmarker.md +135 -114
  191. package/.agent/workflows/plan.md +196 -173
  192. package/.agent/workflows/preview.md +103 -80
  193. package/.agent/workflows/refactor.md +192 -161
  194. package/.agent/workflows/review-ai.md +125 -101
  195. package/.agent/workflows/review.md +141 -116
  196. package/.agent/workflows/session.md +122 -94
  197. package/.agent/workflows/status.md +101 -79
  198. package/.agent/workflows/strengthen-skills.md +164 -138
  199. package/.agent/workflows/super-prompt.md +24 -0
  200. package/.agent/workflows/swarm.md +193 -179
  201. package/.agent/workflows/test.md +211 -189
  202. package/.agent/workflows/tribunal-backend.md +136 -105
  203. package/.agent/workflows/tribunal-database.md +129 -95
  204. package/.agent/workflows/tribunal-frontend.md +140 -96
  205. package/.agent/workflows/tribunal-full.md +131 -100
  206. package/.agent/workflows/tribunal-mobile.md +129 -95
  207. package/.agent/workflows/tribunal-performance.md +136 -110
  208. package/.agent/workflows/tribunal-speed.md +209 -183
  209. package/.agent/workflows/ui-ux-pro-max.md +155 -122
  210. package/README.md +107 -55
  211. package/mcp_config.json +1 -3
  212. package/package.json +94 -94
  213. package/.agent/GEMINI.md +0 -121
  214. package/.agent/skills/doc.md +0 -177
@@ -1,141 +1,145 @@
1
- ---
2
- name: devops-incident-responder
3
- description: Production incident response mastery. MTTR (Mean Time to Recovery) reduction, blameless post-mortems, rapid triaging, halting systemic cascading failures, isolating problematic deployments, and evidence-based forensic analysis. Use when stabilizing broken systems, fighting active production fires, or conducting root-cause post-mortems.
4
- allowed-tools: Read, Write, Edit, Glob, Grep
5
- version: 2.0.0
6
- last-updated: 2026-04-02
7
- applies-to-model: gemini-2.5-pro, claude-3-7-sonnet
8
- ---
9
-
10
- ## Hallucination Traps (Read First)
11
- - ❌ Changing code during an active incident -> ✅ STABILIZE first (rollback, feature flag, traffic shift), investigate AFTER
12
- - ❌ Assigning blame in post-mortems -> ✅ Blameless post-mortems focus on systemic causes, not individual errors
13
- - Skipping the 'what went well' section -> ✅ Understanding what prevented worse outcomes is as valuable as the root cause
14
-
15
- ---
16
-
17
-
18
- # Incident Responder — Production Stabilization Mastery
19
-
20
- ---
21
-
22
- ## 1. The Prime Directive (Stop the Bleeding)
23
-
24
- When an outage is declared (e.g., 502 Bad Gateway across the entire primary cluster), do not ask the developer to check the database logs to figure out why the code crashed.
25
-
26
- **Immediate Action Pipeline:**
27
- 1. **Identify the Trigger:** What changed in the last 15 minutes? (90% of outages are caused by deployments).
28
- 2. **Revert the Change:** Execute the emergency rollback pipeline instantly. Revert the Git commit, swap the Docker tag, or disable the Feature Flag.
29
- 3. **Verify Stabilization:** Ensure metrics return to healthy thresholds.
30
- 4. **Communicate:** "Mitigation complete. Services restored. Root cause investigation underway."
31
-
32
- ---
33
-
34
- ## 2. Isolating Cascading Failures
35
-
36
- A cascading failure occurs when Service A dies, causing Service B to overload with retries, which kills Service B, which kills the database.
37
-
38
- **The Circuit Breaker Protocol:**
39
- If a downstream dependency is dead, sever it immediately to save the rest of the ecosystem.
40
-
41
- ```javascript
42
- // VULNERABLE: Infinite Retry Death Spiral
43
- async function fetchUser(id) {
44
- while(true) {
45
- try { return await api.get(`/user/${id}`); }
46
- catch { await sleep(100); } // Hundreds of containers doing this will execute a DDoSing attack on the API
47
- }
48
- }
49
-
50
- // RESILIENT: Circuit Breaking / Fallbacks
51
- const breaker = new CircuitBreaker(fetchUser, {
52
- errorThresholdPercentage: 50, // If 50% of requests fail...
53
- resetTimeout: 30000 // Open the circuit (stop sending requests) for 30s
54
- });
55
-
56
- breaker.fallback(() => ({ id: "cached-user", status: "degraded" }));
57
- ```
58
-
59
- **Heavy Mitigation Tactics:**
60
- - **Shed Load:** Aggressively drop non-critical traffic (e.g., disable background syncs, temporarily ban aggressive scraping IPs).
61
- - **Scale Out (Band-Aid):** If the memory leak is crashing nodes every 10 minutes, scale the nodes up 3x to buy yourself 30 minutes of runway to find the actual bug.
62
-
63
- ---
64
-
65
- ## 3. The Investigative Triage Routine
66
-
67
- Once the bleeding is stopped (or if you are investigating a non-fatal anomaly), follow the data strictly:
68
-
69
- 1. **Metrics (The "What"):** Look at the Dashboards. Did latency spike? Did CPU pin at 100%? Did Database active connections max out?
70
- 2. **Traces (The "Where"):** Look at OpenTelemetry/Datadog traces. Which specific microservice is the bottleneck?
71
- 3. **Logs (The "Why"):** Query the centralized logs (Splunk/Elastic/CloudWatch) exactly around the timestamp the trace spiked.
72
-
73
- ---
74
-
75
- ## 4. The Blameless Post-Mortem
76
-
77
- Incident response does not end when the system recovers. It ends when the system is architected to survive the same failure tomorrow automatically.
78
-
79
- **A Professional Post-Mortem Must Include:**
80
- 1. **The Timeline:** Chronological factual representation of the event to the minute.
81
- 2. **Root Cause Analysis (The 5 Whys):**
82
- - *Why did the site go down?* DB exhausted connections.
83
- - *Why did it exhaust?* The new background worker didn't pool connections.
84
- - *Why did the worker deploy?* It bypassed CI tests for speed.
85
- 3. **Action Items:** Tangible Jira tickets preventing recurrence (e.g., "Implement PgBouncer connection limits", "Enforce CI checks block on all branches").
86
-
87
- ---
88
-
89
-
90
- ---
91
-
92
-
93
-
94
- AI coding assistants often fall into specific bad habits when dealing with this domain. These are strictly forbidden:
95
-
96
- 1. **Over-engineering:** Proposing complex abstractions or distributed systems when a simpler approach suffices.
97
- 2. **Hallucinated Libraries/Methods:** Using non-existent methods or packages. Always `// VERIFY` or check `package.json` / `requirements.txt`.
98
- 3. **Skipping Edge Cases:** Writing the "happy path" and ignoring error handling, timeouts, or data validation.
99
- 4. **Context Amnesia:** Forgetting the user's constraints and offering generic advice instead of tailored solutions.
100
- 5. **Silent Degradation:** Catching and suppressing errors without logging or re-raising.
101
-
102
- ---
103
-
104
-
105
-
106
- **Slash command: `/review` or `/tribunal-full`**
107
- **Active reviewers: `logic-reviewer` · `security-auditor`**
108
-
109
- ### ❌ Forbidden AI Tropes
110
-
111
- 1. **Blind Assumptions:** Never make an assumption without documenting it clearly with `// VERIFY: [reason]`.
112
- 2. **Silent Degradation:** Catching and suppressing errors without logging or handling.
113
- 3. **Context Amnesia:** Forgetting the user's constraints and offering generic advice instead of tailored solutions.
114
-
115
-
116
-
117
- Review these questions before confirming output:
118
- ```
119
- ✅ Did I rely ONLY on real, verified tools and methods?
120
- ✅ Is this solution appropriately scoped to the user's constraints?
121
- ✅ Did I handle potential failure modes and edge cases?
122
- ✅ Have I avoided generic boilerplate that doesn't add value?
123
- ```
124
-
125
- ### 🛑 Verification-Before-Completion (VBC) Protocol
126
-
127
- **CRITICAL:** You must follow a strict "evidence-based closeout" state machine.
128
- - ❌ **Forbidden:** Declaring a task complete because the output "looks correct."
129
- - ✅ **Required:** You are explicitly forbidden from finalizing any task without providing **concrete evidence** (terminal output, passing tests, compile success, or equivalent proof) that your output works as intended.
130
-
131
-
132
- ## Pre-Flight Checklist
133
- - [ ] Have I reviewed the user's specific constraints and requests?
134
- - [ ] Have I checked the environment for relevant existing implementations?
135
-
136
- ## VBC Protocol (Verification-Before-Completion)
137
- You MUST verify existing code signatures and variables before attempting to modify or call them. No hallucination is permitted.
1
+ ---
2
+ name: devops-incident-responder
3
+ description: Production incident response mastery. MTTR (Mean Time to Recovery) reduction, blameless post-mortems, rapid triaging, halting systemic cascading failures, isolating problematic deployments, and evidence-based forensic analysis. Use when stabilizing broken systems, fighting active production fires, or conducting root-cause post-mortems.
4
+ allowed-tools: Read, Write, Edit, Glob, Grep
5
+ version: 2.0.0
6
+ last-updated: 2026-04-02
7
+ applies-to-model: gemini-2.5-pro, claude-3-7-sonnet
8
+ routing:
9
+ domain: general
10
+ tier: basic
11
+ ---
12
+
13
+ ## Hallucination Traps (Read First)
14
+
15
+ - ❌ Changing code during an active incident -> ✅ STABILIZE first (rollback, feature flag, traffic shift), investigate AFTER
16
+ - ❌ Assigning blame in post-mortems -> ✅ Blameless post-mortems focus on systemic causes, not individual errors
17
+ - ❌ Skipping the 'what went well' section -> ✅ Understanding what prevented worse outcomes is as valuable as the root cause
18
+
19
+ ---
20
+
21
+ # Incident Responder — Production Stabilization Mastery
22
+
23
+ ---
24
+
25
+ ## 1. The Prime Directive (Stop the Bleeding)
26
+
27
+ When an outage is declared (e.g., 502 Bad Gateway across the entire primary cluster), do not ask the developer to check the database logs to figure out why the code crashed.
28
+
29
+ **Immediate Action Pipeline:**
30
+
31
+ 1. **Identify the Trigger:** What changed in the last 15 minutes? (90% of outages are caused by deployments).
32
+ 2. **Revert the Change:** Execute the emergency rollback pipeline instantly. Revert the Git commit, swap the Docker tag, or disable the Feature Flag.
33
+ 3. **Verify Stabilization:** Ensure metrics return to healthy thresholds.
34
+ 4. **Communicate:** "Mitigation complete. Services restored. Root cause investigation underway."
35
+
36
+ ---
37
+
38
+ ## 2. Isolating Cascading Failures
39
+
40
+ A cascading failure occurs when Service A dies, causing Service B to overload with retries, which kills Service B, which kills the database.
41
+
42
+ **The Circuit Breaker Protocol:**
43
+ If a downstream dependency is dead, sever it immediately to save the rest of the ecosystem.
44
+
45
+ ```javascript
46
+ // VULNERABLE: Infinite Retry Death Spiral
47
+ async function fetchUser(id) {
48
+ while (true) {
49
+ try {
50
+ return await api.get(`/user/${id}`);
51
+ } catch {
52
+ await sleep(100);
53
+ } // Hundreds of containers doing this will execute a DDoSing attack on the API
54
+ }
55
+ }
56
+
57
+ // ✅ RESILIENT: Circuit Breaking / Fallbacks
58
+ const breaker = new CircuitBreaker(fetchUser, {
59
+ errorThresholdPercentage: 50, // If 50% of requests fail...
60
+ resetTimeout: 30000, // Open the circuit (stop sending requests) for 30s
61
+ });
138
62
 
63
+ breaker.fallback(() => ({ id: "cached-user", status: "degraded" }));
64
+ ```
65
+
66
+ **Heavy Mitigation Tactics:**
67
+
68
+ - **Shed Load:** Aggressively drop non-critical traffic (e.g., disable background syncs, temporarily ban aggressive scraping IPs).
69
+ - **Scale Out (Band-Aid):** If the memory leak is crashing nodes every 10 minutes, scale the nodes up 3x to buy yourself 30 minutes of runway to find the actual bug.
70
+
71
+ ---
72
+
73
+ ## 3. The Investigative Triage Routine
74
+
75
+ Once the bleeding is stopped (or if you are investigating a non-fatal anomaly), follow the data strictly:
76
+
77
+ 1. **Metrics (The "What"):** Look at the Dashboards. Did latency spike? Did CPU pin at 100%? Did Database active connections max out?
78
+ 2. **Traces (The "Where"):** Look at OpenTelemetry/Datadog traces. Which specific microservice is the bottleneck?
79
+ 3. **Logs (The "Why"):** Query the centralized logs (Splunk/Elastic/CloudWatch) exactly around the timestamp the trace spiked.
80
+
81
+ ---
82
+
83
+ ## 4. The Blameless Post-Mortem
84
+
85
+ Incident response does not end when the system recovers. It ends when the system is architected to survive the same failure tomorrow automatically.
86
+
87
+ **A Professional Post-Mortem Must Include:**
88
+
89
+ 1. **The Timeline:** Chronological factual representation of the event to the minute.
90
+ 2. **Root Cause Analysis (The 5 Whys):**
91
+ - _Why did the site go down?_ DB exhausted connections.
92
+ - _Why did it exhaust?_ The new background worker didn't pool connections.
93
+ - _Why did the worker deploy?_ It bypassed CI tests for speed.
94
+ 3. **Action Items:** Tangible Jira tickets preventing recurrence (e.g., "Implement PgBouncer connection limits", "Enforce CI checks block on all branches").
95
+
96
+ ---
97
+
98
+ ---
99
+
100
+ AI coding assistants often fall into specific bad habits when dealing with this domain. These are strictly forbidden:
101
+
102
+ 1. **Over-engineering:** Proposing complex abstractions or distributed systems when a simpler approach suffices.
103
+ 2. **Hallucinated Libraries/Methods:** Using non-existent methods or packages. Always `// VERIFY` or check `package.json` / `requirements.txt`.
104
+ 3. **Skipping Edge Cases:** Writing the "happy path" and ignoring error handling, timeouts, or data validation.
105
+ 4. **Context Amnesia:** Forgetting the user's constraints and offering generic advice instead of tailored solutions.
106
+ 5. **Silent Degradation:** Catching and suppressing errors without logging or re-raising.
107
+
108
+ ---
109
+
110
+ **Slash command: `/review` or `/tribunal-full`**
111
+ **Active reviewers: `logic-reviewer` · `security-auditor`**
112
+
113
+ ### ❌ Forbidden AI Tropes
114
+
115
+ 1. **Blind Assumptions:** Never make an assumption without documenting it clearly with `// VERIFY: [reason]`.
116
+ 2. **Silent Degradation:** Catching and suppressing errors without logging or handling.
117
+ 3. **Context Amnesia:** Forgetting the user's constraints and offering generic advice instead of tailored solutions.
118
+
119
+ Review these questions before confirming output:
120
+
121
+ ```
122
+ ✅ Did I rely ONLY on real, verified tools and methods?
123
+ ✅ Is this solution appropriately scoped to the user's constraints?
124
+ ✅ Did I handle potential failure modes and edge cases?
125
+ ✅ Have I avoided generic boilerplate that doesn't add value?
126
+ ```
127
+
128
+ ### 🛑 Verification-Before-Completion (VBC) Protocol
129
+
130
+ **CRITICAL:** You must follow a strict "evidence-based closeout" state machine.
131
+
132
+ - ❌ **Forbidden:** Declaring a task complete because the output "looks correct."
133
+ - ✅ **Required:** You are explicitly forbidden from finalizing any task without providing **concrete evidence** (terminal output, passing tests, compile success, or equivalent proof) that your output works as intended.
134
+
135
+ ## Pre-Flight Checklist
136
+
137
+ - [ ] Have I reviewed the user's specific constraints and requests?
138
+ - [ ] Have I checked the environment for relevant existing implementations?
139
+
140
+ ## VBC Protocol (Verification-Before-Completion)
141
+
142
+ You MUST verify existing code signatures and variables before attempting to modify or call them. No hallucination is permitted.
139
143
 
140
144
  ---
141
145
 
@@ -165,6 +169,7 @@ AI coding assistants often fall into specific bad habits when dealing with this
165
169
  ### ✅ Pre-Flight Self-Audit
166
170
 
167
171
  Review these questions before confirming output:
172
+
168
173
  ```
169
174
  ✅ Did I rely ONLY on real, verified tools and methods?
170
175
  ✅ Is this solution appropriately scoped to the user's constraints?
@@ -175,5 +180,6 @@ Review these questions before confirming output:
175
180
  ### 🛑 Verification-Before-Completion (VBC) Protocol
176
181
 
177
182
  **CRITICAL:** You must follow a strict "evidence-based closeout" state machine.
183
+
178
184
  - ❌ **Forbidden:** Declaring a task complete because the output "looks correct."
179
185
  - ✅ **Required:** You are explicitly forbidden from finalizing any task without providing **concrete evidence** (terminal output, passing tests, compile success, or equivalent proof) that your output works as intended.