@rune-kit/rune 2.10.0 → 2.11.0
This diff represents the content of publicly available package versions that have been released to one of the supported registries. The information contained in this diff is provided for informational purposes only and reflects changes between package versions as they appear in their respective public registries.
- package/LICENSE +21 -21
- package/README.md +8 -6
- package/commands/rune.md +168 -168
- package/contexts/dev.md +34 -34
- package/contexts/research.md +43 -43
- package/contexts/review.md +55 -55
- package/extensions/ai-ml/PACK.md +88 -88
- package/extensions/ai-ml/skills/ai-agents.md +172 -172
- package/extensions/ai-ml/skills/code-sandbox.md +187 -187
- package/extensions/ai-ml/skills/deep-research.md +146 -146
- package/extensions/ai-ml/skills/embedding-search.md +66 -66
- package/extensions/ai-ml/skills/fine-tuning-guide.md +74 -74
- package/extensions/ai-ml/skills/llm-architect.md +125 -125
- package/extensions/ai-ml/skills/llm-integration.md +64 -64
- package/extensions/ai-ml/skills/prompt-patterns.md +72 -72
- package/extensions/ai-ml/skills/rag-patterns.md +66 -66
- package/extensions/ai-ml/skills/web-extraction.md +114 -114
- package/extensions/analytics/PACK.md +92 -92
- package/extensions/analytics/skills/ab-testing.md +72 -72
- package/extensions/analytics/skills/dashboard-patterns.md +83 -83
- package/extensions/analytics/skills/data-validation.md +68 -68
- package/extensions/analytics/skills/funnel-analysis.md +81 -81
- package/extensions/analytics/skills/sql-patterns.md +57 -57
- package/extensions/analytics/skills/statistical-analysis.md +79 -79
- package/extensions/analytics/skills/tracking-setup.md +71 -71
- package/extensions/backend/PACK.md +104 -104
- package/extensions/backend/skills/api-patterns.md +84 -84
- package/extensions/backend/skills/async-pipeline.md +193 -193
- package/extensions/backend/skills/auth-patterns.md +97 -97
- package/extensions/backend/skills/background-jobs.md +133 -133
- package/extensions/backend/skills/caching-patterns.md +108 -108
- package/extensions/backend/skills/cli-generation.md +133 -133
- package/extensions/backend/skills/database-patterns.md +87 -87
- package/extensions/backend/skills/middleware-patterns.md +104 -104
- package/extensions/chrome-ext/PACK.md +93 -93
- package/extensions/chrome-ext/skills/cws-preflight.md +143 -143
- package/extensions/chrome-ext/skills/cws-publish.md +104 -104
- package/extensions/chrome-ext/skills/ext-ai-integration.md +251 -251
- package/extensions/chrome-ext/skills/ext-messaging.md +139 -139
- package/extensions/chrome-ext/skills/ext-storage.md +133 -133
- package/extensions/chrome-ext/skills/mv3-scaffold.md +164 -164
- package/extensions/content/PACK.md +96 -96
- package/extensions/content/skills/blog-patterns.md +88 -88
- package/extensions/content/skills/cms-integration.md +131 -131
- package/extensions/content/skills/content-scoring.md +107 -107
- package/extensions/content/skills/i18n.md +83 -83
- package/extensions/content/skills/mdx-authoring.md +137 -137
- package/extensions/content/skills/reference.md +1014 -1014
- package/extensions/content/skills/seo-patterns.md +67 -67
- package/extensions/content/skills/video-repurpose.md +153 -153
- package/extensions/devops/PACK.md +101 -101
- package/extensions/devops/skills/chaos-testing.md +67 -67
- package/extensions/devops/skills/ci-cd.md +75 -75
- package/extensions/devops/skills/docker.md +58 -58
- package/extensions/devops/skills/edge-serverless.md +163 -163
- package/extensions/devops/skills/infra-as-code.md +158 -158
- package/extensions/devops/skills/kubernetes.md +110 -110
- package/extensions/devops/skills/monitoring.md +57 -57
- package/extensions/devops/skills/server-setup.md +64 -64
- package/extensions/devops/skills/ssl-domain.md +42 -42
- package/extensions/ecommerce/PACK.md +116 -116
- package/extensions/ecommerce/skills/cart-system.md +79 -79
- package/extensions/ecommerce/skills/inventory-mgmt.md +102 -102
- package/extensions/ecommerce/skills/order-management.md +126 -126
- package/extensions/ecommerce/skills/payment-integration.md +472 -472
- package/extensions/ecommerce/skills/shopify-dev.md +69 -69
- package/extensions/ecommerce/skills/subscription-billing.md +93 -93
- package/extensions/ecommerce/skills/tax-compliance.md +117 -117
- package/extensions/gamedev/PACK.md +142 -142
- package/extensions/gamedev/skills/asset-pipeline.md +74 -74
- package/extensions/gamedev/skills/audio-system.md +129 -129
- package/extensions/gamedev/skills/camera-system.md +87 -87
- package/extensions/gamedev/skills/ecs.md +98 -98
- package/extensions/gamedev/skills/game-loops.md +72 -72
- package/extensions/gamedev/skills/input-system.md +199 -199
- package/extensions/gamedev/skills/multiplayer.md +180 -180
- package/extensions/gamedev/skills/particles.md +105 -105
- package/extensions/gamedev/skills/physics-engine.md +89 -89
- package/extensions/gamedev/skills/scene-management.md +146 -146
- package/extensions/gamedev/skills/threejs-patterns.md +90 -90
- package/extensions/gamedev/skills/webgl.md +71 -71
- package/extensions/mobile/PACK.md +106 -106
- package/extensions/mobile/skills/app-store-connect.md +152 -152
- package/extensions/mobile/skills/app-store-prep.md +66 -66
- package/extensions/mobile/skills/deep-linking.md +109 -109
- package/extensions/mobile/skills/flutter.md +60 -60
- package/extensions/mobile/skills/ios-build-pipeline.md +142 -142
- package/extensions/mobile/skills/native-bridge.md +66 -66
- package/extensions/mobile/skills/ota-updates.md +97 -97
- package/extensions/mobile/skills/push-notifications.md +111 -111
- package/extensions/mobile/skills/react-native.md +82 -82
- package/extensions/saas/PACK.md +116 -116
- package/extensions/saas/skills/billing-integration.md +200 -200
- package/extensions/saas/skills/feature-flags.md +130 -130
- package/extensions/saas/skills/multi-tenant.md +103 -103
- package/extensions/saas/skills/onboarding-flow.md +139 -139
- package/extensions/saas/skills/subscription-flow.md +95 -95
- package/extensions/saas/skills/team-management.md +144 -144
- package/extensions/security/PACK.md +99 -99
- package/extensions/security/skills/api-security.md +140 -140
- package/extensions/security/skills/compliance.md +68 -68
- package/extensions/security/skills/owasp-audit.md +64 -64
- package/extensions/security/skills/pentest-patterns.md +77 -77
- package/extensions/security/skills/secret-mgmt.md +65 -65
- package/extensions/security/skills/supply-chain.md +65 -65
- package/extensions/trading/PACK.md +80 -80
- package/extensions/trading/skills/chart-components.md +55 -55
- package/extensions/trading/skills/experiment-loop.md +125 -125
- package/extensions/trading/skills/fintech-patterns.md +47 -47
- package/extensions/trading/skills/indicator-library.md +58 -58
- package/extensions/trading/skills/quant-analysis.md +111 -111
- package/extensions/trading/skills/realtime-data.md +58 -58
- package/extensions/trading/skills/trade-logic.md +104 -104
- package/extensions/ui/PACK.md +130 -130
- package/extensions/ui/skills/a11y-audit.md +91 -91
- package/extensions/ui/skills/animation-patterns.md +127 -127
- package/extensions/ui/skills/component-patterns.md +100 -100
- package/extensions/ui/skills/design-decision.md +108 -108
- package/extensions/ui/skills/design-system.md +68 -68
- package/extensions/ui/skills/landing-patterns.md +155 -155
- package/extensions/ui/skills/palette-picker.md +173 -173
- package/extensions/ui/skills/react-health.md +90 -90
- package/extensions/ui/skills/type-system.md +125 -125
- package/extensions/ui/skills/web-vitals.md +153 -153
- package/extensions/zalo/PACK.md +145 -145
- package/extensions/zalo/skills/zalo-oa-mcp.md +317 -317
- package/extensions/zalo/skills/zalo-oa-messaging.md +429 -429
- package/extensions/zalo/skills/zalo-oa-setup.md +236 -236
- package/extensions/zalo/skills/zalo-oa-webhook.md +189 -189
- package/extensions/zalo/skills/zalo-personal-messaging.md +194 -194
- package/extensions/zalo/skills/zalo-personal-setup.md +153 -153
- package/extensions/zalo/skills/zalo-rate-guard.md +219 -219
- package/hooks/auto-format/index.cjs +48 -48
- package/hooks/hooks.json +111 -111
- package/hooks/post-session-reflect/index.cjs +189 -189
- package/hooks/pre-compact/index.cjs +95 -95
- package/hooks/run-hook.cmd +1 -1
- package/hooks/secrets-scan/index.cjs +100 -100
- package/hooks/session-start/index.cjs +71 -71
- package/hooks/typecheck/index.cjs +65 -65
- package/package.json +63 -63
- package/references/ui-pro-max-data/LICENSE-UI-PRO-MAX +21 -21
- package/references/ui-pro-max-data/charts.csv +26 -26
- package/references/ui-pro-max-data/colors.csv +161 -161
- package/references/ui-pro-max-data/styles.csv +68 -68
- package/references/ui-pro-max-data/typography.csv +74 -74
- package/references/ui-pro-max-data/ui-reasoning.csv +162 -162
- package/references/ui-pro-max-data/ux-guidelines.csv +99 -99
- package/skills/adversary/SKILL.md +283 -283
- package/skills/asset-creator/SKILL.md +157 -157
- package/skills/audit/SKILL.md +147 -2
- package/skills/autopsy/SKILL.md +335 -335
- package/skills/brainstorm/SKILL.md +342 -342
- package/skills/browser-pilot/SKILL.md +168 -168
- package/skills/constraint-check/SKILL.md +165 -165
- package/skills/context-engine/SKILL.md +404 -404
- package/skills/cook/SKILL.md +917 -863
- package/skills/db/SKILL.md +273 -273
- package/skills/debug/SKILL.md +465 -465
- package/skills/dependency-doctor/SKILL.md +265 -235
- package/skills/deploy/SKILL.md +274 -231
- package/skills/design/DESIGN-REFERENCE.md +365 -365
- package/skills/design/SKILL.md +589 -589
- package/skills/doc-processor/SKILL.md +254 -254
- package/skills/docs/SKILL.md +374 -374
- package/skills/docs-seeker/SKILL.md +177 -177
- package/skills/fix/SKILL.md +330 -330
- package/skills/git/SKILL.md +339 -339
- package/skills/hallucination-guard/SKILL.md +219 -219
- package/skills/incident/SKILL.md +254 -253
- package/skills/integrity-check/SKILL.md +169 -169
- package/skills/journal/SKILL.md +240 -240
- package/skills/launch/SKILL.md +344 -344
- package/skills/logic-guardian/SKILL.md +251 -251
- package/skills/marketing/SKILL.md +290 -289
- package/skills/mcp-builder/SKILL.md +425 -425
- package/skills/neural-memory/SKILL.md +362 -362
- package/skills/onboard/SKILL.md +404 -403
- package/skills/perf/SKILL.md +346 -346
- package/skills/plan/SKILL.md +433 -428
- package/skills/preflight/SKILL.md +415 -415
- package/skills/problem-solver/SKILL.md +380 -284
- package/skills/rescue/SKILL.md +474 -474
- package/skills/retro/SKILL.md +3 -1
- package/skills/review/SKILL.md +612 -588
- package/skills/review-intake/SKILL.md +249 -249
- package/skills/safeguard/SKILL.md +200 -200
- package/skills/sast/SKILL.md +190 -190
- package/skills/scaffold/SKILL.md +328 -287
- package/skills/scope-guard/SKILL.md +180 -180
- package/skills/scout/SKILL.md +263 -263
- package/skills/sentinel/SKILL.md +382 -381
- package/skills/sentinel-env/SKILL.md +254 -254
- package/skills/sequential-thinking/SKILL.md +234 -234
- package/skills/session-bridge/SKILL.md +543 -543
- package/skills/skill-forge/SKILL.md +581 -581
- package/skills/skill-router/SKILL.md +3 -0
- package/skills/surgeon/SKILL.md +215 -215
- package/skills/team/SKILL.md +556 -537
- package/skills/test/SKILL.md +614 -614
- package/skills/trend-scout/SKILL.md +145 -145
- package/skills/verification/SKILL.md +326 -326
- package/skills/video-creator/SKILL.md +201 -201
- package/skills/watchdog/SKILL.md +168 -168
- package/skills/worktree/SKILL.md +140 -140
|
@@ -1,168 +1,168 @@
|
|
|
1
|
-
---
|
|
2
|
-
name: browser-pilot
|
|
3
|
-
description: Playwright browser automation. Navigates URLs, takes screenshots, checks accessibility tree, interacts with UI elements, and reports findings.
|
|
4
|
-
metadata:
|
|
5
|
-
author: runedev
|
|
6
|
-
version: "0.2.0"
|
|
7
|
-
layer: L3
|
|
8
|
-
model: sonnet
|
|
9
|
-
group: media
|
|
10
|
-
tools: "Read, Bash, Glob, Grep"
|
|
11
|
-
---
|
|
12
|
-
|
|
13
|
-
# browser-pilot
|
|
14
|
-
|
|
15
|
-
## Purpose
|
|
16
|
-
|
|
17
|
-
Browser automation for testing and verification using MCP Playwright tools. Navigates to URLs, captures accessibility snapshots and screenshots, interacts with UI elements (click, type, fill form), and reports findings with visual evidence.
|
|
18
|
-
|
|
19
|
-
## Called By (inbound)
|
|
20
|
-
|
|
21
|
-
- `test` (L2): e2e and visual testing
|
|
22
|
-
- `deploy` (L2): verify live deployment
|
|
23
|
-
- `debug` (L2): capture browser console errors
|
|
24
|
-
- `marketing` (L2): screenshot for assets
|
|
25
|
-
- `launch` (L1): verify live site after deployment
|
|
26
|
-
- `perf` (L2): Lighthouse / Core Web Vitals measurement
|
|
27
|
-
|
|
28
|
-
## Calls (outbound)
|
|
29
|
-
|
|
30
|
-
None — pure L3 utility using Playwright MCP tools.
|
|
31
|
-
|
|
32
|
-
## Executable Instructions
|
|
33
|
-
|
|
34
|
-
### Step 1: Receive Task
|
|
35
|
-
|
|
36
|
-
Accept input from calling skill:
|
|
37
|
-
- `url` — target URL to open
|
|
38
|
-
- `task` — what to do: `screenshot` | `check_elements` | `fill_form` | `test_flow` | `console_errors`
|
|
39
|
-
- `interactions` — optional list of actions (click X, type Y into Z, etc.)
|
|
40
|
-
|
|
41
|
-
### Step 2: Navigate
|
|
42
|
-
|
|
43
|
-
Open the target URL using the Playwright MCP navigate tool:
|
|
44
|
-
|
|
45
|
-
```
|
|
46
|
-
mcp__plugin_playwright_playwright__browser_navigate({ url: "<url>" })
|
|
47
|
-
```
|
|
48
|
-
|
|
49
|
-
Wait for the page to load. If navigation fails (timeout or error), report UNREACHABLE and stop.
|
|
50
|
-
|
|
51
|
-
### Step 3: Snapshot
|
|
52
|
-
|
|
53
|
-
Capture the accessibility tree to understand page structure:
|
|
54
|
-
|
|
55
|
-
```
|
|
56
|
-
mcp__plugin_playwright_playwright__browser_snapshot()
|
|
57
|
-
```
|
|
58
|
-
|
|
59
|
-
Use the snapshot to:
|
|
60
|
-
- Identify interactive elements (buttons, inputs, links)
|
|
61
|
-
- Find specific elements referenced in the task
|
|
62
|
-
- Detect accessibility issues (missing labels, roles)
|
|
63
|
-
|
|
64
|
-
### Step 4: Interact
|
|
65
|
-
|
|
66
|
-
Based on the task, perform interactions using Playwright MCP tools:
|
|
67
|
-
|
|
68
|
-
- **Click**: `mcp__plugin_playwright_playwright__browser_click({ ref: "<ref>", element: "<description>" })`
|
|
69
|
-
- **Type**: `mcp__plugin_playwright_playwright__browser_type({ ref: "<ref>", text: "<value>" })`
|
|
70
|
-
- **Fill form**: `mcp__plugin_playwright_playwright__browser_fill_form({ fields: [...] })`
|
|
71
|
-
- **Navigate back**: `mcp__plugin_playwright_playwright__browser_navigate_back()`
|
|
72
|
-
- **Select option**: `mcp__plugin_playwright_playwright__browser_select_option({ ref: "<ref>", values: [...] })`
|
|
73
|
-
|
|
74
|
-
Limit: max 20 interactions per session. If the task requires more, stop and report partial results.
|
|
75
|
-
|
|
76
|
-
After each interaction, take a new snapshot to verify the result before proceeding.
|
|
77
|
-
|
|
78
|
-
### Step 5: Screenshot
|
|
79
|
-
|
|
80
|
-
Capture visual evidence:
|
|
81
|
-
|
|
82
|
-
```
|
|
83
|
-
mcp__plugin_playwright_playwright__browser_take_screenshot({ type: "png" })
|
|
84
|
-
```
|
|
85
|
-
|
|
86
|
-
For full-page capture (landing pages, long content):
|
|
87
|
-
|
|
88
|
-
```
|
|
89
|
-
mcp__plugin_playwright_playwright__browser_take_screenshot({ type: "png", fullPage: true })
|
|
90
|
-
```
|
|
91
|
-
|
|
92
|
-
Save with a descriptive filename if the `filename` param is supported.
|
|
93
|
-
|
|
94
|
-
### Step 6: Report
|
|
95
|
-
|
|
96
|
-
Compile findings into a structured report:
|
|
97
|
-
|
|
98
|
-
```
|
|
99
|
-
## Browser Report: [url]
|
|
100
|
-
|
|
101
|
-
- **Task**: [task description]
|
|
102
|
-
- **Status**: SUCCESS | PARTIAL | FAILED
|
|
103
|
-
|
|
104
|
-
### Page Info
|
|
105
|
-
- HTTP Status: [status]
|
|
106
|
-
- Load outcome: [loaded | timeout | error]
|
|
107
|
-
|
|
108
|
-
### Accessibility Findings
|
|
109
|
-
- [finding from snapshot — missing labels, broken roles, etc.]
|
|
110
|
-
|
|
111
|
-
### Interaction Log
|
|
112
|
-
- [action taken] → [result: success | element not found | error]
|
|
113
|
-
|
|
114
|
-
### Console Errors
|
|
115
|
-
- [error message — source]
|
|
116
|
-
|
|
117
|
-
### Screenshots
|
|
118
|
-
- [screenshot path or description]
|
|
119
|
-
|
|
120
|
-
### Summary
|
|
121
|
-
- [overall assessment — what works, what failed, any critical issues]
|
|
122
|
-
```
|
|
123
|
-
|
|
124
|
-
### Step 7: Close
|
|
125
|
-
|
|
126
|
-
Always close the browser when done:
|
|
127
|
-
|
|
128
|
-
```
|
|
129
|
-
mcp__plugin_playwright_playwright__browser_close()
|
|
130
|
-
```
|
|
131
|
-
|
|
132
|
-
This step is mandatory even if earlier steps fail. Use a try-finally pattern in your reasoning.
|
|
133
|
-
|
|
134
|
-
## Output Format
|
|
135
|
-
|
|
136
|
-
Structured Browser Report with task status, page info, accessibility findings, interaction log, console errors, screenshots, and summary. See Step 6 Report above for full template.
|
|
137
|
-
|
|
138
|
-
## Constraints
|
|
139
|
-
|
|
140
|
-
1. MUST close browser when done — Step 7 is non-optional even if earlier steps fail
|
|
141
|
-
2. MUST NOT exceed 20 interactions per session
|
|
142
|
-
3. MUST NOT store credentials or sensitive data in interaction logs
|
|
143
|
-
4. MUST take screenshot evidence before reporting visual findings
|
|
144
|
-
|
|
145
|
-
## Sharp Edges
|
|
146
|
-
|
|
147
|
-
Known failure modes for this skill. Check these before declaring done.
|
|
148
|
-
|
|
149
|
-
| Failure Mode | Severity | Mitigation |
|
|
150
|
-
|---|---|---|
|
|
151
|
-
| Not closing browser when done (including on error) | CRITICAL | Constraint 1: Step 7 browser_close() is mandatory — treat as try-finally |
|
|
152
|
-
| Storing credentials or tokens in interaction logs | HIGH | Constraint 3: redact all sensitive values before logging |
|
|
153
|
-
| Exceeding 20 interactions without stopping and reporting partial | MEDIUM | Constraint 2: stop at 20, report what was tested and what remains |
|
|
154
|
-
| Reporting visual findings without screenshot evidence | MEDIUM | Constraint 4: screenshot before reporting — "looks broken" without screenshot is invalid |
|
|
155
|
-
|
|
156
|
-
## Done When
|
|
157
|
-
|
|
158
|
-
- URL navigated successfully (or UNREACHABLE reported)
|
|
159
|
-
- Page snapshot captured for accessibility context
|
|
160
|
-
- All requested interactions completed (or partial with reason if >20)
|
|
161
|
-
- Screenshot taken as visual evidence
|
|
162
|
-
- Console errors captured if task requested them
|
|
163
|
-
- Browser closed (Step 7 executed)
|
|
164
|
-
- Browser Report emitted with status, findings, and screenshot reference
|
|
165
|
-
|
|
166
|
-
## Cost Profile
|
|
167
|
-
|
|
168
|
-
~500-1500 tokens input, ~300-800 tokens output. Sonnet for interaction logic.
|
|
1
|
+
---
|
|
2
|
+
name: browser-pilot
|
|
3
|
+
description: Playwright browser automation. Navigates URLs, takes screenshots, checks accessibility tree, interacts with UI elements, and reports findings.
|
|
4
|
+
metadata:
|
|
5
|
+
author: runedev
|
|
6
|
+
version: "0.2.0"
|
|
7
|
+
layer: L3
|
|
8
|
+
model: sonnet
|
|
9
|
+
group: media
|
|
10
|
+
tools: "Read, Bash, Glob, Grep"
|
|
11
|
+
---
|
|
12
|
+
|
|
13
|
+
# browser-pilot
|
|
14
|
+
|
|
15
|
+
## Purpose
|
|
16
|
+
|
|
17
|
+
Browser automation for testing and verification using MCP Playwright tools. Navigates to URLs, captures accessibility snapshots and screenshots, interacts with UI elements (click, type, fill form), and reports findings with visual evidence.
|
|
18
|
+
|
|
19
|
+
## Called By (inbound)
|
|
20
|
+
|
|
21
|
+
- `test` (L2): e2e and visual testing
|
|
22
|
+
- `deploy` (L2): verify live deployment
|
|
23
|
+
- `debug` (L2): capture browser console errors
|
|
24
|
+
- `marketing` (L2): screenshot for assets
|
|
25
|
+
- `launch` (L1): verify live site after deployment
|
|
26
|
+
- `perf` (L2): Lighthouse / Core Web Vitals measurement
|
|
27
|
+
|
|
28
|
+
## Calls (outbound)
|
|
29
|
+
|
|
30
|
+
None — pure L3 utility using Playwright MCP tools.
|
|
31
|
+
|
|
32
|
+
## Executable Instructions
|
|
33
|
+
|
|
34
|
+
### Step 1: Receive Task
|
|
35
|
+
|
|
36
|
+
Accept input from calling skill:
|
|
37
|
+
- `url` — target URL to open
|
|
38
|
+
- `task` — what to do: `screenshot` | `check_elements` | `fill_form` | `test_flow` | `console_errors`
|
|
39
|
+
- `interactions` — optional list of actions (click X, type Y into Z, etc.)
|
|
40
|
+
|
|
41
|
+
### Step 2: Navigate
|
|
42
|
+
|
|
43
|
+
Open the target URL using the Playwright MCP navigate tool:
|
|
44
|
+
|
|
45
|
+
```
|
|
46
|
+
mcp__plugin_playwright_playwright__browser_navigate({ url: "<url>" })
|
|
47
|
+
```
|
|
48
|
+
|
|
49
|
+
Wait for the page to load. If navigation fails (timeout or error), report UNREACHABLE and stop.
|
|
50
|
+
|
|
51
|
+
### Step 3: Snapshot
|
|
52
|
+
|
|
53
|
+
Capture the accessibility tree to understand page structure:
|
|
54
|
+
|
|
55
|
+
```
|
|
56
|
+
mcp__plugin_playwright_playwright__browser_snapshot()
|
|
57
|
+
```
|
|
58
|
+
|
|
59
|
+
Use the snapshot to:
|
|
60
|
+
- Identify interactive elements (buttons, inputs, links)
|
|
61
|
+
- Find specific elements referenced in the task
|
|
62
|
+
- Detect accessibility issues (missing labels, roles)
|
|
63
|
+
|
|
64
|
+
### Step 4: Interact
|
|
65
|
+
|
|
66
|
+
Based on the task, perform interactions using Playwright MCP tools:
|
|
67
|
+
|
|
68
|
+
- **Click**: `mcp__plugin_playwright_playwright__browser_click({ ref: "<ref>", element: "<description>" })`
|
|
69
|
+
- **Type**: `mcp__plugin_playwright_playwright__browser_type({ ref: "<ref>", text: "<value>" })`
|
|
70
|
+
- **Fill form**: `mcp__plugin_playwright_playwright__browser_fill_form({ fields: [...] })`
|
|
71
|
+
- **Navigate back**: `mcp__plugin_playwright_playwright__browser_navigate_back()`
|
|
72
|
+
- **Select option**: `mcp__plugin_playwright_playwright__browser_select_option({ ref: "<ref>", values: [...] })`
|
|
73
|
+
|
|
74
|
+
Limit: max 20 interactions per session. If the task requires more, stop and report partial results.
|
|
75
|
+
|
|
76
|
+
After each interaction, take a new snapshot to verify the result before proceeding.
|
|
77
|
+
|
|
78
|
+
### Step 5: Screenshot
|
|
79
|
+
|
|
80
|
+
Capture visual evidence:
|
|
81
|
+
|
|
82
|
+
```
|
|
83
|
+
mcp__plugin_playwright_playwright__browser_take_screenshot({ type: "png" })
|
|
84
|
+
```
|
|
85
|
+
|
|
86
|
+
For full-page capture (landing pages, long content):
|
|
87
|
+
|
|
88
|
+
```
|
|
89
|
+
mcp__plugin_playwright_playwright__browser_take_screenshot({ type: "png", fullPage: true })
|
|
90
|
+
```
|
|
91
|
+
|
|
92
|
+
Save with a descriptive filename if the `filename` param is supported.
|
|
93
|
+
|
|
94
|
+
### Step 6: Report
|
|
95
|
+
|
|
96
|
+
Compile findings into a structured report:
|
|
97
|
+
|
|
98
|
+
```
|
|
99
|
+
## Browser Report: [url]
|
|
100
|
+
|
|
101
|
+
- **Task**: [task description]
|
|
102
|
+
- **Status**: SUCCESS | PARTIAL | FAILED
|
|
103
|
+
|
|
104
|
+
### Page Info
|
|
105
|
+
- HTTP Status: [status]
|
|
106
|
+
- Load outcome: [loaded | timeout | error]
|
|
107
|
+
|
|
108
|
+
### Accessibility Findings
|
|
109
|
+
- [finding from snapshot — missing labels, broken roles, etc.]
|
|
110
|
+
|
|
111
|
+
### Interaction Log
|
|
112
|
+
- [action taken] → [result: success | element not found | error]
|
|
113
|
+
|
|
114
|
+
### Console Errors
|
|
115
|
+
- [error message — source]
|
|
116
|
+
|
|
117
|
+
### Screenshots
|
|
118
|
+
- [screenshot path or description]
|
|
119
|
+
|
|
120
|
+
### Summary
|
|
121
|
+
- [overall assessment — what works, what failed, any critical issues]
|
|
122
|
+
```
|
|
123
|
+
|
|
124
|
+
### Step 7: Close
|
|
125
|
+
|
|
126
|
+
Always close the browser when done:
|
|
127
|
+
|
|
128
|
+
```
|
|
129
|
+
mcp__plugin_playwright_playwright__browser_close()
|
|
130
|
+
```
|
|
131
|
+
|
|
132
|
+
This step is mandatory even if earlier steps fail. Use a try-finally pattern in your reasoning.
|
|
133
|
+
|
|
134
|
+
## Output Format
|
|
135
|
+
|
|
136
|
+
Structured Browser Report with task status, page info, accessibility findings, interaction log, console errors, screenshots, and summary. See Step 6 Report above for full template.
|
|
137
|
+
|
|
138
|
+
## Constraints
|
|
139
|
+
|
|
140
|
+
1. MUST close browser when done — Step 7 is non-optional even if earlier steps fail
|
|
141
|
+
2. MUST NOT exceed 20 interactions per session
|
|
142
|
+
3. MUST NOT store credentials or sensitive data in interaction logs
|
|
143
|
+
4. MUST take screenshot evidence before reporting visual findings
|
|
144
|
+
|
|
145
|
+
## Sharp Edges
|
|
146
|
+
|
|
147
|
+
Known failure modes for this skill. Check these before declaring done.
|
|
148
|
+
|
|
149
|
+
| Failure Mode | Severity | Mitigation |
|
|
150
|
+
|---|---|---|
|
|
151
|
+
| Not closing browser when done (including on error) | CRITICAL | Constraint 1: Step 7 browser_close() is mandatory — treat as try-finally |
|
|
152
|
+
| Storing credentials or tokens in interaction logs | HIGH | Constraint 3: redact all sensitive values before logging |
|
|
153
|
+
| Exceeding 20 interactions without stopping and reporting partial | MEDIUM | Constraint 2: stop at 20, report what was tested and what remains |
|
|
154
|
+
| Reporting visual findings without screenshot evidence | MEDIUM | Constraint 4: screenshot before reporting — "looks broken" without screenshot is invalid |
|
|
155
|
+
|
|
156
|
+
## Done When
|
|
157
|
+
|
|
158
|
+
- URL navigated successfully (or UNREACHABLE reported)
|
|
159
|
+
- Page snapshot captured for accessibility context
|
|
160
|
+
- All requested interactions completed (or partial with reason if >20)
|
|
161
|
+
- Screenshot taken as visual evidence
|
|
162
|
+
- Console errors captured if task requested them
|
|
163
|
+
- Browser closed (Step 7 executed)
|
|
164
|
+
- Browser Report emitted with status, findings, and screenshot reference
|
|
165
|
+
|
|
166
|
+
## Cost Profile
|
|
167
|
+
|
|
168
|
+
~500-1500 tokens input, ~300-800 tokens output. Sonnet for interaction logic.
|
|
@@ -1,165 +1,165 @@
|
|
|
1
|
-
---
|
|
2
|
-
name: constraint-check
|
|
3
|
-
description: "Meta-validator for HARD-GATEs. Verifies that skills' mandatory constraints were followed during a workflow. Called by cook, team, and audit to audit discipline compliance."
|
|
4
|
-
user-invocable: false
|
|
5
|
-
metadata:
|
|
6
|
-
author: runedev
|
|
7
|
-
version: "1.1.0"
|
|
8
|
-
layer: L3
|
|
9
|
-
model: haiku
|
|
10
|
-
group: validation
|
|
11
|
-
tools: "Read, Glob, Grep"
|
|
12
|
-
---
|
|
13
|
-
|
|
14
|
-
# constraint-check
|
|
15
|
-
|
|
16
|
-
## Purpose
|
|
17
|
-
|
|
18
|
-
The internal affairs department for Rune skills. Checks whether HARD-GATEs and mandatory constraints were actually followed during a workflow — not just claimed to be followed. Reads the constraint definitions from skill files and audits the conversation trail for compliance.
|
|
19
|
-
|
|
20
|
-
While `completion-gate` checks if claims have evidence, `constraint-check` checks if the PROCESS was followed. Did you actually write tests before code? Did you actually get plan approval? Did you actually run sentinel?
|
|
21
|
-
|
|
22
|
-
## Triggers
|
|
23
|
-
|
|
24
|
-
- Called by `cook` (L1) at end of workflow as discipline audit
|
|
25
|
-
- Called by `team` (L1) to verify stream agents followed constraints
|
|
26
|
-
- Called by `audit` (L2) during quality dimension assessment
|
|
27
|
-
- `/rune constraint-check` — manual audit of current session
|
|
28
|
-
|
|
29
|
-
## Calls (outbound)
|
|
30
|
-
|
|
31
|
-
None — pure read-only validator.
|
|
32
|
-
|
|
33
|
-
## Called By (inbound)
|
|
34
|
-
|
|
35
|
-
- `cook` (L1): end-of-workflow discipline audit
|
|
36
|
-
- `team` (L1): verify stream agent compliance
|
|
37
|
-
- `audit` (L2): quality dimension
|
|
38
|
-
- User: manual session audit
|
|
39
|
-
|
|
40
|
-
## Execution
|
|
41
|
-
|
|
42
|
-
### Step 1 — Identify Active Skills
|
|
43
|
-
|
|
44
|
-
Parse the conversation/workflow to identify which skills were invoked:
|
|
45
|
-
|
|
46
|
-
```
|
|
47
|
-
Extract from context:
|
|
48
|
-
- Skills invoked via Skill tool (exact list)
|
|
49
|
-
- Skills referenced in agent narrative
|
|
50
|
-
- Phase progression (cook phases completed)
|
|
51
|
-
```
|
|
52
|
-
|
|
53
|
-
### Step 2 — Load Constraint Definitions
|
|
54
|
-
|
|
55
|
-
For each invoked skill, extract HARD-GATEs and numbered constraints:
|
|
56
|
-
|
|
57
|
-
```
|
|
58
|
-
For each skill in invoked_skills:
|
|
59
|
-
Read: skills/<skill>/SKILL.md
|
|
60
|
-
Extract:
|
|
61
|
-
- <HARD-GATE> blocks → mandatory, violation = BLOCK
|
|
62
|
-
- ## Constraints numbered list → required, violation = WARN
|
|
63
|
-
- ## Mesh Gates table → required gates
|
|
64
|
-
```
|
|
65
|
-
|
|
66
|
-
### Step 3 — Audit Compliance
|
|
67
|
-
|
|
68
|
-
Check each constraint against the conversation evidence:
|
|
69
|
-
|
|
70
|
-
| Constraint Type | How to Verify | Evidence Source |
|
|
71
|
-
|---|---|---|
|
|
72
|
-
| "MUST write tests BEFORE code" | Test file Write/Edit timestamps before implementation Write/Edit | Tool call ordering |
|
|
73
|
-
| "MUST get user approval" | User message containing "go"/"yes"/"proceed" after plan | Conversation history |
|
|
74
|
-
| "MUST run verification" | Bash command with test/lint/build output | Tool call results |
|
|
75
|
-
| "MUST show actual output" | Stdout captured in agent response | Agent messages |
|
|
76
|
-
| "MUST NOT modify files outside scope" | Git diff files vs plan file list | Git + plan comparison |
|
|
77
|
-
| "Iron Law: delete code before test" | No implementation code exists before test creation | Tool call ordering |
|
|
78
|
-
|
|
79
|
-
### Step 4 — Classify Violations
|
|
80
|
-
|
|
81
|
-
| Violation Type | Severity | Meaning |
|
|
82
|
-
|---------------|----------|---------|
|
|
83
|
-
| HARD-GATE violation | BLOCK | Skill says this is non-negotiable |
|
|
84
|
-
| Constraint violation | WARN | Skill says this is required but not fatal |
|
|
85
|
-
| Best practice skip | INFO | Recommended but optional |
|
|
86
|
-
|
|
87
|
-
### Step 5 — Report
|
|
88
|
-
|
|
89
|
-
```
|
|
90
|
-
## Constraint Check Report
|
|
91
|
-
- **Status**: COMPLIANT | VIOLATIONS_FOUND | CRITICAL_VIOLATION
|
|
92
|
-
- **Skills Audited**: [count]
|
|
93
|
-
- **Constraints Checked**: [count]
|
|
94
|
-
- **Violations**: [count by severity]
|
|
95
|
-
|
|
96
|
-
### HARD-GATE Violations (BLOCK)
|
|
97
|
-
- [skill:test] Iron Law: implementation code written at tool_call #12 BEFORE test file created at #15
|
|
98
|
-
- [skill:cook] Plan Gate: Phase 4 started without user approval message
|
|
99
|
-
|
|
100
|
-
### Constraint Violations (WARN)
|
|
101
|
-
- [skill:verification] Constraint 2: "All tests pass" claimed at message #20 without stdout evidence
|
|
102
|
-
- [skill:sentinel] Constraint 3: files scanned list not included in report
|
|
103
|
-
|
|
104
|
-
### Compliance Summary
|
|
105
|
-
| Skill | HARD-GATEs | Constraints | Status |
|
|
106
|
-
|-------|-----------|-------------|--------|
|
|
107
|
-
| cook | 3/3 ✓ | 6/7 (1 WARN) | WARN |
|
|
108
|
-
| test | 0/1 ✗ | 8/9 (1 WARN) | BLOCK |
|
|
109
|
-
| verification | 1/1 ✓ | 4/6 (2 WARN) | WARN |
|
|
110
|
-
| sentinel | 1/1 ✓ | 7/7 ✓ | PASS |
|
|
111
|
-
|
|
112
|
-
### Remediation
|
|
113
|
-
- BLOCK: test Iron Law — delete implementation, restart with test-first
|
|
114
|
-
- WARN: verification — re-run and capture stdout
|
|
115
|
-
```
|
|
116
|
-
|
|
117
|
-
## Constraint Catalog (Quick Reference)
|
|
118
|
-
|
|
119
|
-
Key HARD-GATEs across skills that constraint-check audits:
|
|
120
|
-
|
|
121
|
-
| Skill | HARD-GATE | Check Method |
|
|
122
|
-
|---|---|---|
|
|
123
|
-
| test | Tests BEFORE code (Iron Law) | Tool call ordering |
|
|
124
|
-
| cook | Scout before plan, plan before code | Phase progression |
|
|
125
|
-
| plan | Every code phase has test entry | Plan content |
|
|
126
|
-
| verification | Evidence for every claim | Stdout capture |
|
|
127
|
-
| sentinel | BLOCK = halt pipeline | No commit after BLOCK |
|
|
128
|
-
| preflight | BLOCK = halt pipeline | No commit after BLOCK |
|
|
129
|
-
| debug | No code changes during debug | No Write/Edit in debug |
|
|
130
|
-
| debug | 3-fix escalation | Fix attempt counter |
|
|
131
|
-
| brainstorm | No implementation before approval | User message check |
|
|
132
|
-
|
|
133
|
-
## Output Format
|
|
134
|
-
|
|
135
|
-
Constraint Check Report with status (COMPLIANT/VIOLATIONS_FOUND/CRITICAL_VIOLATION), HARD-GATE violations, constraint violations, compliance summary table, and remediation steps. See Step 5 Report above for full template.
|
|
136
|
-
|
|
137
|
-
## Constraints
|
|
138
|
-
|
|
139
|
-
1. MUST check all HARD-GATEs for every invoked skill — not just the ones that seem relevant
|
|
140
|
-
2. MUST use tool call ordering (not agent narrative) to verify temporal constraints
|
|
141
|
-
3. MUST distinguish HARD-GATE violations (BLOCK) from constraint violations (WARN)
|
|
142
|
-
4. MUST report specific evidence for each violation — not just "violated"
|
|
143
|
-
5. MUST NOT accept agent's self-report as compliance evidence — check independently
|
|
144
|
-
|
|
145
|
-
## Sharp Edges
|
|
146
|
-
|
|
147
|
-
| Failure Mode | Severity | Mitigation |
|
|
148
|
-
|---|---|---|
|
|
149
|
-
| Agent self-reports compliance and constraint-check trusts it | CRITICAL | Constraint 5: check tool calls independently, not agent narrative |
|
|
150
|
-
| Only checking cook constraints, missing test/sentinel/etc | HIGH | Constraint 1: audit ALL invoked skills, not just the orchestrator |
|
|
151
|
-
| Temporal check wrong (tool calls reordered in context) | MEDIUM | Use tool call sequence numbers, not message ordering |
|
|
152
|
-
| Too strict on optional steps (INFO treated as BLOCK) | LOW | Step 4 classification: only HARD-GATE = BLOCK, constraints = WARN |
|
|
153
|
-
|
|
154
|
-
## Done When
|
|
155
|
-
|
|
156
|
-
- All invoked skills identified from context
|
|
157
|
-
- HARD-GATEs and constraints extracted from each skill's SKILL.md
|
|
158
|
-
- Each constraint checked against conversation evidence
|
|
159
|
-
- Violations classified as BLOCK/WARN/INFO
|
|
160
|
-
- Compliance summary table emitted per skill
|
|
161
|
-
- Remediation steps listed for each violation
|
|
162
|
-
|
|
163
|
-
## Cost Profile
|
|
164
|
-
|
|
165
|
-
~1000-2000 tokens input, ~500-1000 tokens output. Haiku for speed — reads skill files and checks tool call ordering.
|
|
1
|
+
---
|
|
2
|
+
name: constraint-check
|
|
3
|
+
description: "Meta-validator for HARD-GATEs. Verifies that skills' mandatory constraints were followed during a workflow. Called by cook, team, and audit to audit discipline compliance."
|
|
4
|
+
user-invocable: false
|
|
5
|
+
metadata:
|
|
6
|
+
author: runedev
|
|
7
|
+
version: "1.1.0"
|
|
8
|
+
layer: L3
|
|
9
|
+
model: haiku
|
|
10
|
+
group: validation
|
|
11
|
+
tools: "Read, Glob, Grep"
|
|
12
|
+
---
|
|
13
|
+
|
|
14
|
+
# constraint-check
|
|
15
|
+
|
|
16
|
+
## Purpose
|
|
17
|
+
|
|
18
|
+
The internal affairs department for Rune skills. Checks whether HARD-GATEs and mandatory constraints were actually followed during a workflow — not just claimed to be followed. Reads the constraint definitions from skill files and audits the conversation trail for compliance.
|
|
19
|
+
|
|
20
|
+
While `completion-gate` checks if claims have evidence, `constraint-check` checks if the PROCESS was followed. Did you actually write tests before code? Did you actually get plan approval? Did you actually run sentinel?
|
|
21
|
+
|
|
22
|
+
## Triggers
|
|
23
|
+
|
|
24
|
+
- Called by `cook` (L1) at end of workflow as discipline audit
|
|
25
|
+
- Called by `team` (L1) to verify stream agents followed constraints
|
|
26
|
+
- Called by `audit` (L2) during quality dimension assessment
|
|
27
|
+
- `/rune constraint-check` — manual audit of current session
|
|
28
|
+
|
|
29
|
+
## Calls (outbound)
|
|
30
|
+
|
|
31
|
+
None — pure read-only validator.
|
|
32
|
+
|
|
33
|
+
## Called By (inbound)
|
|
34
|
+
|
|
35
|
+
- `cook` (L1): end-of-workflow discipline audit
|
|
36
|
+
- `team` (L1): verify stream agent compliance
|
|
37
|
+
- `audit` (L2): quality dimension
|
|
38
|
+
- User: manual session audit
|
|
39
|
+
|
|
40
|
+
## Execution
|
|
41
|
+
|
|
42
|
+
### Step 1 — Identify Active Skills
|
|
43
|
+
|
|
44
|
+
Parse the conversation/workflow to identify which skills were invoked:
|
|
45
|
+
|
|
46
|
+
```
|
|
47
|
+
Extract from context:
|
|
48
|
+
- Skills invoked via Skill tool (exact list)
|
|
49
|
+
- Skills referenced in agent narrative
|
|
50
|
+
- Phase progression (cook phases completed)
|
|
51
|
+
```
|
|
52
|
+
|
|
53
|
+
### Step 2 — Load Constraint Definitions
|
|
54
|
+
|
|
55
|
+
For each invoked skill, extract HARD-GATEs and numbered constraints:
|
|
56
|
+
|
|
57
|
+
```
|
|
58
|
+
For each skill in invoked_skills:
|
|
59
|
+
Read: skills/<skill>/SKILL.md
|
|
60
|
+
Extract:
|
|
61
|
+
- <HARD-GATE> blocks → mandatory, violation = BLOCK
|
|
62
|
+
- ## Constraints numbered list → required, violation = WARN
|
|
63
|
+
- ## Mesh Gates table → required gates
|
|
64
|
+
```
|
|
65
|
+
|
|
66
|
+
### Step 3 — Audit Compliance
|
|
67
|
+
|
|
68
|
+
Check each constraint against the conversation evidence:
|
|
69
|
+
|
|
70
|
+
| Constraint Type | How to Verify | Evidence Source |
|
|
71
|
+
|---|---|---|
|
|
72
|
+
| "MUST write tests BEFORE code" | Test file Write/Edit timestamps before implementation Write/Edit | Tool call ordering |
|
|
73
|
+
| "MUST get user approval" | User message containing "go"/"yes"/"proceed" after plan | Conversation history |
|
|
74
|
+
| "MUST run verification" | Bash command with test/lint/build output | Tool call results |
|
|
75
|
+
| "MUST show actual output" | Stdout captured in agent response | Agent messages |
|
|
76
|
+
| "MUST NOT modify files outside scope" | Git diff files vs plan file list | Git + plan comparison |
|
|
77
|
+
| "Iron Law: delete code before test" | No implementation code exists before test creation | Tool call ordering |
|
|
78
|
+
|
|
79
|
+
### Step 4 — Classify Violations
|
|
80
|
+
|
|
81
|
+
| Violation Type | Severity | Meaning |
|
|
82
|
+
|---------------|----------|---------|
|
|
83
|
+
| HARD-GATE violation | BLOCK | Skill says this is non-negotiable |
|
|
84
|
+
| Constraint violation | WARN | Skill says this is required but not fatal |
|
|
85
|
+
| Best practice skip | INFO | Recommended but optional |
|
|
86
|
+
|
|
87
|
+
### Step 5 — Report
|
|
88
|
+
|
|
89
|
+
```
|
|
90
|
+
## Constraint Check Report
|
|
91
|
+
- **Status**: COMPLIANT | VIOLATIONS_FOUND | CRITICAL_VIOLATION
|
|
92
|
+
- **Skills Audited**: [count]
|
|
93
|
+
- **Constraints Checked**: [count]
|
|
94
|
+
- **Violations**: [count by severity]
|
|
95
|
+
|
|
96
|
+
### HARD-GATE Violations (BLOCK)
|
|
97
|
+
- [skill:test] Iron Law: implementation code written at tool_call #12 BEFORE test file created at #15
|
|
98
|
+
- [skill:cook] Plan Gate: Phase 4 started without user approval message
|
|
99
|
+
|
|
100
|
+
### Constraint Violations (WARN)
|
|
101
|
+
- [skill:verification] Constraint 2: "All tests pass" claimed at message #20 without stdout evidence
|
|
102
|
+
- [skill:sentinel] Constraint 3: files scanned list not included in report
|
|
103
|
+
|
|
104
|
+
### Compliance Summary
|
|
105
|
+
| Skill | HARD-GATEs | Constraints | Status |
|
|
106
|
+
|-------|-----------|-------------|--------|
|
|
107
|
+
| cook | 3/3 ✓ | 6/7 (1 WARN) | WARN |
|
|
108
|
+
| test | 0/1 ✗ | 8/9 (1 WARN) | BLOCK |
|
|
109
|
+
| verification | 1/1 ✓ | 4/6 (2 WARN) | WARN |
|
|
110
|
+
| sentinel | 1/1 ✓ | 7/7 ✓ | PASS |
|
|
111
|
+
|
|
112
|
+
### Remediation
|
|
113
|
+
- BLOCK: test Iron Law — delete implementation, restart with test-first
|
|
114
|
+
- WARN: verification — re-run and capture stdout
|
|
115
|
+
```
|
|
116
|
+
|
|
117
|
+
## Constraint Catalog (Quick Reference)
|
|
118
|
+
|
|
119
|
+
Key HARD-GATEs across skills that constraint-check audits:
|
|
120
|
+
|
|
121
|
+
| Skill | HARD-GATE | Check Method |
|
|
122
|
+
|---|---|---|
|
|
123
|
+
| test | Tests BEFORE code (Iron Law) | Tool call ordering |
|
|
124
|
+
| cook | Scout before plan, plan before code | Phase progression |
|
|
125
|
+
| plan | Every code phase has test entry | Plan content |
|
|
126
|
+
| verification | Evidence for every claim | Stdout capture |
|
|
127
|
+
| sentinel | BLOCK = halt pipeline | No commit after BLOCK |
|
|
128
|
+
| preflight | BLOCK = halt pipeline | No commit after BLOCK |
|
|
129
|
+
| debug | No code changes during debug | No Write/Edit in debug |
|
|
130
|
+
| debug | 3-fix escalation | Fix attempt counter |
|
|
131
|
+
| brainstorm | No implementation before approval | User message check |
|
|
132
|
+
|
|
133
|
+
## Output Format
|
|
134
|
+
|
|
135
|
+
Constraint Check Report with status (COMPLIANT/VIOLATIONS_FOUND/CRITICAL_VIOLATION), HARD-GATE violations, constraint violations, compliance summary table, and remediation steps. See Step 5 Report above for full template.
|
|
136
|
+
|
|
137
|
+
## Constraints
|
|
138
|
+
|
|
139
|
+
1. MUST check all HARD-GATEs for every invoked skill — not just the ones that seem relevant
|
|
140
|
+
2. MUST use tool call ordering (not agent narrative) to verify temporal constraints
|
|
141
|
+
3. MUST distinguish HARD-GATE violations (BLOCK) from constraint violations (WARN)
|
|
142
|
+
4. MUST report specific evidence for each violation — not just "violated"
|
|
143
|
+
5. MUST NOT accept agent's self-report as compliance evidence — check independently
|
|
144
|
+
|
|
145
|
+
## Sharp Edges
|
|
146
|
+
|
|
147
|
+
| Failure Mode | Severity | Mitigation |
|
|
148
|
+
|---|---|---|
|
|
149
|
+
| Agent self-reports compliance and constraint-check trusts it | CRITICAL | Constraint 5: check tool calls independently, not agent narrative |
|
|
150
|
+
| Only checking cook constraints, missing test/sentinel/etc | HIGH | Constraint 1: audit ALL invoked skills, not just the orchestrator |
|
|
151
|
+
| Temporal check wrong (tool calls reordered in context) | MEDIUM | Use tool call sequence numbers, not message ordering |
|
|
152
|
+
| Too strict on optional steps (INFO treated as BLOCK) | LOW | Step 4 classification: only HARD-GATE = BLOCK, constraints = WARN |
|
|
153
|
+
|
|
154
|
+
## Done When
|
|
155
|
+
|
|
156
|
+
- All invoked skills identified from context
|
|
157
|
+
- HARD-GATEs and constraints extracted from each skill's SKILL.md
|
|
158
|
+
- Each constraint checked against conversation evidence
|
|
159
|
+
- Violations classified as BLOCK/WARN/INFO
|
|
160
|
+
- Compliance summary table emitted per skill
|
|
161
|
+
- Remediation steps listed for each violation
|
|
162
|
+
|
|
163
|
+
## Cost Profile
|
|
164
|
+
|
|
165
|
+
~1000-2000 tokens input, ~500-1000 tokens output. Haiku for speed — reads skill files and checks tool call ordering.
|