@selesai/code 0.9.6 → 0.9.9
This diff represents the content of publicly available package versions that have been released to one of the supported registries. The information contained in this diff is provided for informational purposes only and reflects changes between package versions as they appear in their respective public registries.
- package/CHANGELOG.md +18 -0
- package/dist/cli/credential-print.js +4 -4
- package/dist/cli.js +0 -0
- package/dist/core/agent-session-runtime.js +1 -4
- package/dist/core/agent-session.d.ts +2 -0
- package/dist/core/agent-session.js +26 -0
- package/dist/core/compaction/branch-summarization.js +30 -30
- package/dist/core/compaction/compaction.js +81 -81
- package/dist/core/compaction/utils.js +2 -2
- package/dist/core/export-html/template.css +1066 -1066
- package/dist/core/export-html/template.html +55 -55
- package/dist/core/export-html/template.js +1864 -1864
- package/dist/core/export-html/vendor/highlight.min.js +1212 -1212
- package/dist/core/export-html/vendor/marked.min.js +78 -78
- package/dist/core/handoff.d.ts +11 -0
- package/dist/core/handoff.js +49 -0
- package/dist/core/messages.js +7 -7
- package/dist/extensions/context-compaction-reminder.test.ts +82 -82
- package/dist/extensions/context-compaction-reminder.ts +28 -28
- package/dist/extensions/handoff-new.ts +14 -75
- package/dist/extensions/pi-intercom/LICENSE +21 -21
- package/dist/extensions/pi-intercom/broker/client.test.ts +83 -83
- package/dist/extensions/pi-intercom/broker/extension.test.ts +387 -387
- package/dist/extensions/pi-intercom/broker/framing.test.ts +114 -114
- package/dist/extensions/pi-intercom/broker/paths.test.ts +153 -153
- package/dist/extensions/pi-intercom/broker/paths.ts +134 -134
- package/dist/extensions/pi-intercom/broker/runtime-claim.test.ts +34 -34
- package/dist/extensions/pi-intercom/broker/runtime-claim.ts +21 -21
- package/dist/extensions/pi-intercom/cwd.test.ts +40 -40
- package/dist/extensions/pi-intercom/cwd.ts +31 -31
- package/dist/extensions/pi-intercom/extension-api.ts +44 -44
- package/dist/extensions/pi-intercom/format-context.test.ts +31 -31
- package/dist/extensions/pi-intercom/format-context.ts +32 -32
- package/dist/extensions/pi-intercom/test/overlay-width.test.ts +66 -66
- package/dist/extensions/pi-intercom/ui/compose.ts +143 -143
- package/dist/extensions/pi-intercom/ui/session-list.ts +166 -166
- package/dist/extensions/pi-powerline-footer/bash-mode/shell-session.ts +286 -286
- package/dist/extensions/pi-powerline-footer/bash-mode/transcript.ts +108 -108
- package/dist/extensions/pi-powerline-footer/separators.ts +57 -57
- package/dist/extensions/pi-powerline-footer/tests/session-usage.test.ts +47 -47
- package/dist/extensions/pi-powerline-footer/tests/tps.test.ts +39 -39
- package/dist/extensions/pi-powerline-footer/theme.example.json +24 -24
- package/dist/extensions/pi-powerline-footer/theme.json +12 -12
- package/dist/extensions/pi-powerline-footer/tps.ts +345 -345
- package/dist/extensions/pi-rewind-hook/README.md +245 -245
- package/dist/extensions/pi-rewind-hook/index.ts +1445 -1445
- package/dist/extensions/pi-rewind-hook/package.json +29 -29
- package/dist/extensions/pi-subagents/install.mjs +0 -0
- package/dist/extensions/pi-web-agent/package.json +31 -31
- package/dist/extensions/pi-web-agent/src/backends/config.ts +205 -205
- package/dist/extensions/pi-web-agent/src/backends/doctor.ts +136 -136
- package/dist/extensions/pi-web-agent/src/backends/factory.ts +152 -152
- package/dist/extensions/pi-web-agent/src/backends/settings-reader.ts +25 -25
- package/dist/extensions/pi-web-agent/src/cache/ttl-cache.ts +28 -28
- package/dist/extensions/pi-web-agent/src/changelog-notice.ts +136 -136
- package/dist/extensions/pi-web-agent/src/commands/web-agent-config.ts +946 -946
- package/dist/extensions/pi-web-agent/src/extension.ts +126 -126
- package/dist/extensions/pi-web-agent/src/extract/readability.ts +118 -118
- package/dist/extensions/pi-web-agent/src/fetch/browser-resolution.ts +199 -199
- package/dist/extensions/pi-web-agent/src/fetch/firecrawl-fetch.ts +100 -100
- package/dist/extensions/pi-web-agent/src/fetch/headless-fetch.ts +117 -117
- package/dist/extensions/pi-web-agent/src/fetch/http-fetch.ts +67 -67
- package/dist/extensions/pi-web-agent/src/orchestration/answer-synthesizer.ts +60 -60
- package/dist/extensions/pi-web-agent/src/orchestration/candidate-selector.ts +50 -50
- package/dist/extensions/pi-web-agent/src/orchestration/direct-url.ts +52 -52
- package/dist/extensions/pi-web-agent/src/orchestration/evidence-quality.ts +105 -105
- package/dist/extensions/pi-web-agent/src/orchestration/evidence-ranker.ts +45 -45
- package/dist/extensions/pi-web-agent/src/orchestration/index.ts +28 -28
- package/dist/extensions/pi-web-agent/src/orchestration/query-planner.ts +47 -47
- package/dist/extensions/pi-web-agent/src/orchestration/research-orchestrator.ts +376 -376
- package/dist/extensions/pi-web-agent/src/orchestration/research-types.ts +64 -64
- package/dist/extensions/pi-web-agent/src/orchestration/research-worker.ts +181 -181
- package/dist/extensions/pi-web-agent/src/orchestration/source-profile.ts +101 -101
- package/dist/extensions/pi-web-agent/src/orchestration/stop-decider.ts +81 -81
- package/dist/extensions/pi-web-agent/src/presentation/config-store.ts +210 -210
- package/dist/extensions/pi-web-agent/src/presentation/config.ts +75 -75
- package/dist/extensions/pi-web-agent/src/presentation/explore-presentation.ts +61 -61
- package/dist/extensions/pi-web-agent/src/presentation/fetch-presentation.ts +54 -54
- package/dist/extensions/pi-web-agent/src/presentation/search-presentation.ts +41 -41
- package/dist/extensions/pi-web-agent/src/presentation/select-view.ts +20 -20
- package/dist/extensions/pi-web-agent/src/presentation/types.ts +63 -63
- package/dist/extensions/pi-web-agent/src/search/brave.ts +114 -114
- package/dist/extensions/pi-web-agent/src/search/duckduckgo.ts +72 -72
- package/dist/extensions/pi-web-agent/src/search/searxng.ts +96 -96
- package/dist/extensions/pi-web-agent/src/tools/web-explore.ts +71 -71
- package/dist/extensions/pi-web-agent/src/tools/web-fetch-headless.ts +31 -31
- package/dist/extensions/pi-web-agent/src/tools/web-fetch.ts +31 -31
- package/dist/extensions/pi-web-agent/src/tools/web-search.ts +157 -157
- package/dist/extensions/pi-web-agent/src/types.ts +88 -88
- package/dist/extensions/ponytail/package.json +8 -8
- package/dist/extensions/question/batch.ts +103 -103
- package/dist/extensions/question/constants.ts +30 -30
- package/dist/extensions/question/helpers.ts +58 -58
- package/dist/extensions/question/navigation.ts +14 -14
- package/dist/extensions/question/package.json +19 -19
- package/dist/extensions/question/schemas.ts +43 -43
- package/dist/extensions/question/selection-mode.ts +46 -46
- package/dist/extensions/question/shortcuts.ts +44 -44
- package/dist/extensions/question/types.ts +132 -132
- package/dist/extensions/test-resolve-hook-impl.mjs +6 -6
- package/dist/extensions/test-resolve-hook.mjs +6 -6
- package/dist/extensions/web-agent-onboarding.ts +222 -222
- package/dist/extensions/workflow/package.json +17 -17
- package/dist/modes/rpc/rpc-mode.js +24 -0
- package/dist/modes/rpc/rpc-types.d.ts +19 -0
- package/dist/package-manager-cli.js +71 -71
- package/dist/rpc-entry.js +0 -0
- package/dist/skills/agent-browser/SKILL.md +52 -52
- package/dist/skills/batch-grill-me/SKILL.md +19 -19
- package/dist/skills/grill-me/SKILL.md +10 -10
- package/dist/skills/handoff/SKILL.md +16 -16
- package/dist/skills/handoff-text/SKILL.md +14 -14
- package/dist/skills/implanger/SKILL.md +69 -69
- package/dist/skills/improve-codebase/REFERENCE.md +78 -78
- package/dist/skills/improve-codebase/SKILL.md +178 -178
- package/dist/skills/planger/SKILL.md +166 -166
- package/dist/skills/ponytail/SKILL.md +116 -116
- package/dist/skills/ponytail-audit/SKILL.md +42 -42
- package/dist/skills/ponytail-debt/SKILL.md +45 -45
- package/dist/skills/ponytail-gain/SKILL.md +51 -51
- package/dist/skills/ponytail-help/SKILL.md +70 -70
- package/dist/skills/ponytail-review/SKILL.md +58 -58
- package/dist/skills/selesai-handoff/SKILL.md +20 -20
- package/dist/skills/workflow-creation/SKILL.md +73 -73
- package/dist/themes/powerline-footer/theme.json +33 -33
- package/docs/compaction.md +396 -396
- package/docs/containerization.md +111 -111
- package/docs/development.md +71 -71
- package/docs/docs.json +164 -164
- package/docs/environment-variables.md +86 -86
- package/docs/index.md +83 -83
- package/docs/json.md +82 -82
- package/docs/models.md +502 -502
- package/docs/packages.md +227 -227
- package/docs/plans/subagent-delegation/phase-0-correctness.md +264 -264
- package/docs/plans/subagent-delegation/phase-1-behavioral-contract.md +485 -485
- package/docs/plans/subagent-delegation/phase-2-context-controls.md +281 -281
- package/docs/plans/subagent-delegation/phase-3-advisory-routing.md +361 -361
- package/docs/plans/subagent-delegation/phase-4-optional-enforcement.md +380 -380
- package/docs/prompt-templates.md +95 -95
- package/docs/providers.md +293 -293
- package/docs/sdk.md +1143 -1143
- package/docs/security.md +59 -59
- package/docs/session-format.md +414 -414
- package/docs/sessions.md +145 -145
- package/docs/shared-host-extensions.md +109 -109
- package/docs/shell-aliases.md +13 -13
- package/docs/skills.md +231 -231
- package/docs/terminal-setup.md +142 -142
- package/docs/termux.md +127 -127
- package/docs/themes.md +295 -295
- package/docs/tmux.md +63 -63
- package/docs/tui.md +927 -927
- package/docs/windows.md +17 -17
- package/examples/README.md +25 -25
- package/examples/extensions/README.md +211 -211
- package/examples/extensions/auto-commit-on-exit.ts +49 -49
- package/examples/extensions/bash-spawn-hook.ts +30 -30
- package/examples/extensions/bookmark.ts +50 -50
- package/examples/extensions/border-status-editor.ts +150 -150
- package/examples/extensions/built-in-tool-renderer.ts +249 -249
- package/examples/extensions/claude-rules.ts +86 -86
- package/examples/extensions/commands.ts +72 -72
- package/examples/extensions/confirm-destructive.ts +59 -59
- package/examples/extensions/custom-compaction.ts +130 -130
- package/examples/extensions/custom-footer.ts +64 -64
- package/examples/extensions/custom-header.ts +73 -73
- package/examples/extensions/custom-provider-anthropic/index.ts +610 -610
- package/examples/extensions/custom-provider-anthropic/package-lock.json +24 -24
- package/examples/extensions/custom-provider-anthropic/package.json +19 -19
- package/examples/extensions/custom-provider-gitlab-duo/index.ts +404 -404
- package/examples/extensions/custom-provider-gitlab-duo/package.json +16 -16
- package/examples/extensions/custom-provider-gitlab-duo/test.ts +82 -82
- package/examples/extensions/dirty-repo-guard.ts +56 -56
- package/examples/extensions/doom-overlay/README.md +46 -46
- package/examples/extensions/doom-overlay/doom/build/doom.js +21 -21
- package/examples/extensions/doom-overlay/doom/build/doom.wasm +0 -0
- package/examples/extensions/doom-overlay/doom/build.sh +152 -152
- package/examples/extensions/doom-overlay/doom/doomgeneric_pi.c +72 -72
- package/examples/extensions/doom-overlay/doom-component.ts +132 -132
- package/examples/extensions/doom-overlay/doom-engine.ts +173 -173
- package/examples/extensions/doom-overlay/doom-keys.ts +104 -104
- package/examples/extensions/doom-overlay/index.ts +74 -74
- package/examples/extensions/doom-overlay/wad-finder.ts +51 -51
- package/examples/extensions/dynamic-resources/SKILL.md +8 -8
- package/examples/extensions/dynamic-resources/dynamic.json +79 -79
- package/examples/extensions/dynamic-resources/dynamic.md +5 -5
- package/examples/extensions/dynamic-resources/index.ts +15 -15
- package/examples/extensions/dynamic-tools.ts +74 -74
- package/examples/extensions/event-bus.ts +43 -43
- package/examples/extensions/file-trigger.ts +41 -41
- package/examples/extensions/git-checkpoint.ts +53 -53
- package/examples/extensions/git-merge-and-resolve.ts +115 -115
- package/examples/extensions/github-issue-autocomplete.ts +185 -185
- package/examples/extensions/gondolin/index.ts +531 -531
- package/examples/extensions/gondolin/package-lock.json +185 -185
- package/examples/extensions/gondolin/package.json +19 -19
- package/examples/extensions/handoff.ts +199 -199
- package/examples/extensions/hello.ts +26 -26
- package/examples/extensions/hidden-thinking-label.ts +53 -53
- package/examples/extensions/inline-bash.ts +94 -94
- package/examples/extensions/input-transform-streaming.ts +39 -39
- package/examples/extensions/input-transform.ts +43 -43
- package/examples/extensions/interactive-shell.ts +196 -196
- package/examples/extensions/mac-system-theme.ts +47 -47
- package/examples/extensions/message-renderer.ts +59 -59
- package/examples/extensions/minimal-mode.ts +426 -426
- package/examples/extensions/modal-editor.ts +85 -85
- package/examples/extensions/model-status.ts +31 -31
- package/examples/extensions/notify.ts +55 -55
- package/examples/extensions/overlay-qa-tests.ts +1450 -1450
- package/examples/extensions/overlay-test.ts +153 -153
- package/examples/extensions/permission-gate.ts +34 -34
- package/examples/extensions/pirate.ts +47 -47
- package/examples/extensions/plan-mode/README.md +66 -66
- package/examples/extensions/plan-mode/index.ts +390 -390
- package/examples/extensions/plan-mode/utils.ts +168 -168
- package/examples/extensions/preset.ts +436 -436
- package/examples/extensions/project-trust.ts +64 -64
- package/examples/extensions/prompt-customizer.ts +97 -97
- package/examples/extensions/protected-paths.ts +30 -30
- package/examples/extensions/provider-payload.ts +18 -18
- package/examples/extensions/qna.ts +122 -122
- package/examples/extensions/question.ts +285 -285
- package/examples/extensions/questionnaire.ts +448 -448
- package/examples/extensions/rainbow-editor.ts +88 -88
- package/examples/extensions/reload-runtime.ts +37 -37
- package/examples/extensions/rpc-demo.ts +118 -118
- package/examples/extensions/sandbox/index.ts +321 -321
- package/examples/extensions/sandbox/package-lock.json +92 -92
- package/examples/extensions/sandbox/package.json +19 -19
- package/examples/extensions/send-user-message.ts +97 -97
- package/examples/extensions/session-name.ts +27 -27
- package/examples/extensions/shutdown-command.ts +63 -63
- package/examples/extensions/snake.ts +343 -343
- package/examples/extensions/space-invaders.ts +560 -560
- package/examples/extensions/ssh.ts +220 -220
- package/examples/extensions/status-line.ts +32 -32
- package/examples/extensions/structured-output.ts +65 -65
- package/examples/extensions/subagent/README.md +175 -175
- package/examples/extensions/subagent/agents/planner.md +37 -37
- package/examples/extensions/subagent/agents/reviewer.md +35 -35
- package/examples/extensions/subagent/agents/scout.md +50 -50
- package/examples/extensions/subagent/agents/worker.md +24 -24
- package/examples/extensions/subagent/agents.ts +126 -126
- package/examples/extensions/subagent/index.ts +1015 -1015
- package/examples/extensions/subagent/prompts/implement-and-review.md +10 -10
- package/examples/extensions/subagent/prompts/implement.md +10 -10
- package/examples/extensions/subagent/prompts/scout-and-plan.md +9 -9
- package/examples/extensions/summarize.ts +209 -209
- package/examples/extensions/system-prompt-header.ts +17 -17
- package/examples/extensions/tic-tac-toe.ts +1008 -1008
- package/examples/extensions/timed-confirm.ts +70 -70
- package/examples/extensions/titlebar-spinner.ts +58 -58
- package/examples/extensions/todo.ts +297 -297
- package/examples/extensions/tool-override.ts +144 -144
- package/examples/extensions/tools.ts +146 -146
- package/examples/extensions/trigger-compact.ts +50 -50
- package/examples/extensions/truncated-tool.ts +195 -195
- package/examples/extensions/widget-placement.ts +9 -9
- package/examples/extensions/with-deps/index.ts +32 -32
- package/examples/extensions/with-deps/package-lock.json +31 -31
- package/examples/extensions/with-deps/package.json +22 -22
- package/examples/extensions/working-indicator.ts +123 -123
- package/examples/extensions/working-message-test.ts +25 -25
- package/examples/rpc-extension-ui.ts +632 -632
- package/examples/sdk/01-minimal.ts +26 -26
- package/examples/sdk/02-custom-model.ts +53 -53
- package/examples/sdk/03-custom-prompt.ts +75 -75
- package/examples/sdk/04-skills.ts +55 -55
- package/examples/sdk/05-tools.ts +48 -48
- package/examples/sdk/06-extensions.ts +99 -99
- package/examples/sdk/07-context-files.ts +47 -47
- package/examples/sdk/08-prompt-templates.ts +51 -51
- package/examples/sdk/09-api-keys-and-oauth.ts +52 -52
- package/examples/sdk/10-settings.ts +53 -53
- package/examples/sdk/11-sessions.ts +52 -52
- package/examples/sdk/12-full-control.ts +79 -79
- package/examples/sdk/13-session-runtime.ts +67 -67
- package/examples/sdk/README.md +144 -144
- package/package.json +1 -1
|
@@ -1,45 +1,45 @@
|
|
|
1
|
-
---
|
|
2
|
-
name: ponytail-debt
|
|
3
|
-
description: >
|
|
4
|
-
Harvest every `ponytail:` comment in the codebase into a debt ledger, so the
|
|
5
|
-
deliberate shortcuts and deferrals ponytail leaves behind get tracked instead
|
|
6
|
-
of rotting into "later means never". Use when the user says "ponytail debt",
|
|
7
|
-
"/ponytail-debt", "what did ponytail defer", "list the shortcuts", "ponytail
|
|
8
|
-
ledger", or "what did we mark to do later". One-shot report, changes nothing.
|
|
9
|
-
disable-model-invocation: true
|
|
10
|
-
---
|
|
11
|
-
|
|
12
|
-
Every deliberate ponytail shortcut is marked with a `ponytail:` comment naming
|
|
13
|
-
its ceiling and upgrade path. This collects them into one ledger so a deferral
|
|
14
|
-
can't quietly become permanent.
|
|
15
|
-
|
|
16
|
-
## Scan
|
|
17
|
-
|
|
18
|
-
Grep the repo for comment markers, skipping `node_modules`, `.git`, and build
|
|
19
|
-
output:
|
|
20
|
-
|
|
21
|
-
`grep -rnE '(#|//) ?ponytail:' .` (add other comment prefixes if your stack uses them)
|
|
22
|
-
|
|
23
|
-
Each hit is one ledger row. The comment prefix keeps prose that merely mentions
|
|
24
|
-
the convention out of the ledger.
|
|
25
|
-
|
|
26
|
-
## Output
|
|
27
|
-
|
|
28
|
-
One row per marker, grouped by file:
|
|
29
|
-
|
|
30
|
-
`<file>:<line>, <what was simplified>. ceiling: <the limit named>. upgrade: <the trigger to revisit>.`
|
|
31
|
-
|
|
32
|
-
The convention is `ponytail: <ceiling>, <upgrade path>`, so pull the ceiling
|
|
33
|
-
and the trigger straight from the comment. Want an owner per row too? add
|
|
34
|
-
`git blame -L<line>,<line>`.
|
|
35
|
-
|
|
36
|
-
Flag the rot risk: any `ponytail:` comment that names no upgrade path or
|
|
37
|
-
trigger gets a `no-trigger` tag, those are the ones that silently rot.
|
|
38
|
-
|
|
39
|
-
End with `<N> markers, <M> with no trigger.` Nothing found: `No ponytail: debt. Clean ledger.`
|
|
40
|
-
|
|
41
|
-
## Boundaries
|
|
42
|
-
|
|
43
|
-
Reads and reports only, changes nothing. To persist it, ask and it writes the
|
|
44
|
-
ledger to a file (e.g. `PONYTAIL-DEBT.md`). One-shot. "stop ponytail-debt" or
|
|
45
|
-
"normal mode" to revert.
|
|
1
|
+
---
|
|
2
|
+
name: ponytail-debt
|
|
3
|
+
description: >
|
|
4
|
+
Harvest every `ponytail:` comment in the codebase into a debt ledger, so the
|
|
5
|
+
deliberate shortcuts and deferrals ponytail leaves behind get tracked instead
|
|
6
|
+
of rotting into "later means never". Use when the user says "ponytail debt",
|
|
7
|
+
"/ponytail-debt", "what did ponytail defer", "list the shortcuts", "ponytail
|
|
8
|
+
ledger", or "what did we mark to do later". One-shot report, changes nothing.
|
|
9
|
+
disable-model-invocation: true
|
|
10
|
+
---
|
|
11
|
+
|
|
12
|
+
Every deliberate ponytail shortcut is marked with a `ponytail:` comment naming
|
|
13
|
+
its ceiling and upgrade path. This collects them into one ledger so a deferral
|
|
14
|
+
can't quietly become permanent.
|
|
15
|
+
|
|
16
|
+
## Scan
|
|
17
|
+
|
|
18
|
+
Grep the repo for comment markers, skipping `node_modules`, `.git`, and build
|
|
19
|
+
output:
|
|
20
|
+
|
|
21
|
+
`grep -rnE '(#|//) ?ponytail:' .` (add other comment prefixes if your stack uses them)
|
|
22
|
+
|
|
23
|
+
Each hit is one ledger row. The comment prefix keeps prose that merely mentions
|
|
24
|
+
the convention out of the ledger.
|
|
25
|
+
|
|
26
|
+
## Output
|
|
27
|
+
|
|
28
|
+
One row per marker, grouped by file:
|
|
29
|
+
|
|
30
|
+
`<file>:<line>, <what was simplified>. ceiling: <the limit named>. upgrade: <the trigger to revisit>.`
|
|
31
|
+
|
|
32
|
+
The convention is `ponytail: <ceiling>, <upgrade path>`, so pull the ceiling
|
|
33
|
+
and the trigger straight from the comment. Want an owner per row too? add
|
|
34
|
+
`git blame -L<line>,<line>`.
|
|
35
|
+
|
|
36
|
+
Flag the rot risk: any `ponytail:` comment that names no upgrade path or
|
|
37
|
+
trigger gets a `no-trigger` tag, those are the ones that silently rot.
|
|
38
|
+
|
|
39
|
+
End with `<N> markers, <M> with no trigger.` Nothing found: `No ponytail: debt. Clean ledger.`
|
|
40
|
+
|
|
41
|
+
## Boundaries
|
|
42
|
+
|
|
43
|
+
Reads and reports only, changes nothing. To persist it, ask and it writes the
|
|
44
|
+
ledger to a file (e.g. `PONYTAIL-DEBT.md`). One-shot. "stop ponytail-debt" or
|
|
45
|
+
"normal mode" to revert.
|
|
@@ -1,51 +1,51 @@
|
|
|
1
|
-
---
|
|
2
|
-
name: ponytail-gain
|
|
3
|
-
description: >
|
|
4
|
-
Show ponytail's measured impact as a compact scoreboard: less code, less
|
|
5
|
-
cost, more speed, from the benchmark medians. One-shot display, not a
|
|
6
|
-
persistent mode, and not a per-repo number. Trigger: /ponytail-gain,
|
|
7
|
-
"ponytail gain", "what does ponytail save", "show ponytail impact",
|
|
8
|
-
"ponytail scoreboard".
|
|
9
|
-
disable-model-invocation: true
|
|
10
|
-
---
|
|
11
|
-
|
|
12
|
-
# Ponytail Gain
|
|
13
|
-
|
|
14
|
-
Display this scoreboard when invoked. One-shot: do NOT change mode, write flag
|
|
15
|
-
files, or persist anything.
|
|
16
|
-
|
|
17
|
-
The figures are the published benchmark medians (5 everyday tasks: email
|
|
18
|
-
validator, debounce, CSV sum, countdown timer, rate limiter; three models:
|
|
19
|
-
Haiku, Sonnet, Opus). They are measured, not computed from the current repo.
|
|
20
|
-
Source: `benchmarks/` and the README.
|
|
21
|
-
|
|
22
|
-
## Scoreboard
|
|
23
|
-
|
|
24
|
-
Render plain ASCII bars. The bar length shows the measured range; the label
|
|
25
|
-
carries the exact figure:
|
|
26
|
-
|
|
27
|
-
```
|
|
28
|
-
ponytail gain benchmark median · 5 tasks · 3 models
|
|
29
|
-
|
|
30
|
-
Lines of code no-skill ████████████████████ 100%
|
|
31
|
-
ponytail ██▌················· 6–20% ▼ 80–94%
|
|
32
|
-
Cost no-skill ████████████████████ 100%
|
|
33
|
-
ponytail █████▌·············· 23–53% ▼ 47–77%
|
|
34
|
-
Speed ponytail ▸ 3–6× faster
|
|
35
|
-
|
|
36
|
-
This repo: /ponytail-debt (shortcuts you deferred)
|
|
37
|
-
/ponytail-audit (what's still cuttable)
|
|
38
|
-
```
|
|
39
|
-
|
|
40
|
-
## Honesty boundary
|
|
41
|
-
|
|
42
|
-
These are benchmark medians, not this repo. NEVER print a per-repo savings
|
|
43
|
-
number ("you saved X lines/tokens here"): the unbuilt version was never
|
|
44
|
-
written, so there is no real baseline to subtract from in a live repo. The
|
|
45
|
-
only real per-repo figures come from `/ponytail-debt` (a counted ledger), and
|
|
46
|
-
this card points there instead of inventing one.
|
|
47
|
-
|
|
48
|
-
## Boundaries
|
|
49
|
-
|
|
50
|
-
One-shot display. Edits nothing, changes no mode.
|
|
51
|
-
"stop ponytail" or "normal mode": revert.
|
|
1
|
+
---
|
|
2
|
+
name: ponytail-gain
|
|
3
|
+
description: >
|
|
4
|
+
Show ponytail's measured impact as a compact scoreboard: less code, less
|
|
5
|
+
cost, more speed, from the benchmark medians. One-shot display, not a
|
|
6
|
+
persistent mode, and not a per-repo number. Trigger: /ponytail-gain,
|
|
7
|
+
"ponytail gain", "what does ponytail save", "show ponytail impact",
|
|
8
|
+
"ponytail scoreboard".
|
|
9
|
+
disable-model-invocation: true
|
|
10
|
+
---
|
|
11
|
+
|
|
12
|
+
# Ponytail Gain
|
|
13
|
+
|
|
14
|
+
Display this scoreboard when invoked. One-shot: do NOT change mode, write flag
|
|
15
|
+
files, or persist anything.
|
|
16
|
+
|
|
17
|
+
The figures are the published benchmark medians (5 everyday tasks: email
|
|
18
|
+
validator, debounce, CSV sum, countdown timer, rate limiter; three models:
|
|
19
|
+
Haiku, Sonnet, Opus). They are measured, not computed from the current repo.
|
|
20
|
+
Source: `benchmarks/` and the README.
|
|
21
|
+
|
|
22
|
+
## Scoreboard
|
|
23
|
+
|
|
24
|
+
Render plain ASCII bars. The bar length shows the measured range; the label
|
|
25
|
+
carries the exact figure:
|
|
26
|
+
|
|
27
|
+
```
|
|
28
|
+
ponytail gain benchmark median · 5 tasks · 3 models
|
|
29
|
+
|
|
30
|
+
Lines of code no-skill ████████████████████ 100%
|
|
31
|
+
ponytail ██▌················· 6–20% ▼ 80–94%
|
|
32
|
+
Cost no-skill ████████████████████ 100%
|
|
33
|
+
ponytail █████▌·············· 23–53% ▼ 47–77%
|
|
34
|
+
Speed ponytail ▸ 3–6× faster
|
|
35
|
+
|
|
36
|
+
This repo: /ponytail-debt (shortcuts you deferred)
|
|
37
|
+
/ponytail-audit (what's still cuttable)
|
|
38
|
+
```
|
|
39
|
+
|
|
40
|
+
## Honesty boundary
|
|
41
|
+
|
|
42
|
+
These are benchmark medians, not this repo. NEVER print a per-repo savings
|
|
43
|
+
number ("you saved X lines/tokens here"): the unbuilt version was never
|
|
44
|
+
written, so there is no real baseline to subtract from in a live repo. The
|
|
45
|
+
only real per-repo figures come from `/ponytail-debt` (a counted ledger), and
|
|
46
|
+
this card points there instead of inventing one.
|
|
47
|
+
|
|
48
|
+
## Boundaries
|
|
49
|
+
|
|
50
|
+
One-shot display. Edits nothing, changes no mode.
|
|
51
|
+
"stop ponytail" or "normal mode": revert.
|
|
@@ -1,70 +1,70 @@
|
|
|
1
|
-
---
|
|
2
|
-
name: ponytail-help
|
|
3
|
-
description: >
|
|
4
|
-
Quick-reference card for all ponytail modes, skills, and commands.
|
|
5
|
-
One-shot display, not a persistent mode. Trigger: /ponytail-help,
|
|
6
|
-
"ponytail help", "what ponytail commands", "how do I use ponytail".
|
|
7
|
-
disable-model-invocation: true
|
|
8
|
-
---
|
|
9
|
-
|
|
10
|
-
# Ponytail Help
|
|
11
|
-
|
|
12
|
-
Display this reference card when invoked. One-shot, do NOT change mode,
|
|
13
|
-
write flag files, or persist anything.
|
|
14
|
-
|
|
15
|
-
## Levels
|
|
16
|
-
|
|
17
|
-
| Level | Trigger | What change |
|
|
18
|
-
|-------|---------|-------------|
|
|
19
|
-
| **Lite** | `/ponytail lite` | Build what's asked, name the lazier alternative in one line. |
|
|
20
|
-
| **Full** | `/ponytail` | The ladder enforced: YAGNI → stdlib → native → one line → minimum. Default. |
|
|
21
|
-
| **Ultra** | `/ponytail ultra` | YAGNI extremist. Deletion before addition. Challenges requirements before building. |
|
|
22
|
-
|
|
23
|
-
Level sticks until changed or session end.
|
|
24
|
-
|
|
25
|
-
## Skills
|
|
26
|
-
|
|
27
|
-
| Skill | Trigger | What it does |
|
|
28
|
-
|-------|---------|--------------|
|
|
29
|
-
| **ponytail** | `/ponytail` | Lazy mode itself. Simplest solution that works. |
|
|
30
|
-
| **ponytail-review** | `/ponytail-review` | Over-engineering review: `L42: yagni: factory, one product. Inline.` |
|
|
31
|
-
| **ponytail-gain** | `/ponytail-gain` | Measured-impact scoreboard: less code, less cost, more speed. |
|
|
32
|
-
| **ponytail-help** | `/ponytail-help` | This card. |
|
|
33
|
-
|
|
34
|
-
Codex uses `@ponytail`, `@ponytail-review`, and `@ponytail-help`; Claude Code
|
|
35
|
-
and OpenCode use the slash-command forms above (OpenCode ships `/ponytail` and
|
|
36
|
-
`/ponytail-review`).
|
|
37
|
-
|
|
38
|
-
## Deactivate
|
|
39
|
-
|
|
40
|
-
Say "stop ponytail" or "normal mode". Resume anytime with `/ponytail`.
|
|
41
|
-
`/ponytail off` also works.
|
|
42
|
-
|
|
43
|
-
## Configure Default Mode
|
|
44
|
-
|
|
45
|
-
Default mode = `full`, auto-active every session. Change it:
|
|
46
|
-
|
|
47
|
-
**Environment variable** (highest priority):
|
|
48
|
-
```bash
|
|
49
|
-
export PONYTAIL_DEFAULT_MODE=ultra
|
|
50
|
-
```
|
|
51
|
-
|
|
52
|
-
**Config file** (`~/.config/ponytail/config.json`, Windows: `%APPDATA%\ponytail\config.json`):
|
|
53
|
-
```json
|
|
54
|
-
{ "defaultMode": "lite" }
|
|
55
|
-
```
|
|
56
|
-
|
|
57
|
-
Set `"off"` to disable auto-activation on session start, activate manually
|
|
58
|
-
with `/ponytail` when wanted.
|
|
59
|
-
|
|
60
|
-
Resolution: env var > config file > `full`.
|
|
61
|
-
|
|
62
|
-
## Update
|
|
63
|
-
|
|
64
|
-
Enable auto-update once: open `/plugin`, go to Marketplaces, pick ponytail, Enable auto-update. Claude Code then pulls new versions at startup (run `/reload-plugins` when it prompts). Manual refresh: `/plugin marketplace update ponytail` then `/reload-plugins`.
|
|
65
|
-
|
|
66
|
-
If `/plugin` is not recognized, your Claude Code is out of date. Update it (`npm install -g @anthropic-ai/claude-code@latest`, or `brew upgrade claude-code`) and restart. Other hosts use their own update flow.
|
|
67
|
-
|
|
68
|
-
## More
|
|
69
|
-
|
|
70
|
-
Full docs + examples: https://github.com/DietrichGebert/ponytail
|
|
1
|
+
---
|
|
2
|
+
name: ponytail-help
|
|
3
|
+
description: >
|
|
4
|
+
Quick-reference card for all ponytail modes, skills, and commands.
|
|
5
|
+
One-shot display, not a persistent mode. Trigger: /ponytail-help,
|
|
6
|
+
"ponytail help", "what ponytail commands", "how do I use ponytail".
|
|
7
|
+
disable-model-invocation: true
|
|
8
|
+
---
|
|
9
|
+
|
|
10
|
+
# Ponytail Help
|
|
11
|
+
|
|
12
|
+
Display this reference card when invoked. One-shot, do NOT change mode,
|
|
13
|
+
write flag files, or persist anything.
|
|
14
|
+
|
|
15
|
+
## Levels
|
|
16
|
+
|
|
17
|
+
| Level | Trigger | What change |
|
|
18
|
+
|-------|---------|-------------|
|
|
19
|
+
| **Lite** | `/ponytail lite` | Build what's asked, name the lazier alternative in one line. |
|
|
20
|
+
| **Full** | `/ponytail` | The ladder enforced: YAGNI → stdlib → native → one line → minimum. Default. |
|
|
21
|
+
| **Ultra** | `/ponytail ultra` | YAGNI extremist. Deletion before addition. Challenges requirements before building. |
|
|
22
|
+
|
|
23
|
+
Level sticks until changed or session end.
|
|
24
|
+
|
|
25
|
+
## Skills
|
|
26
|
+
|
|
27
|
+
| Skill | Trigger | What it does |
|
|
28
|
+
|-------|---------|--------------|
|
|
29
|
+
| **ponytail** | `/ponytail` | Lazy mode itself. Simplest solution that works. |
|
|
30
|
+
| **ponytail-review** | `/ponytail-review` | Over-engineering review: `L42: yagni: factory, one product. Inline.` |
|
|
31
|
+
| **ponytail-gain** | `/ponytail-gain` | Measured-impact scoreboard: less code, less cost, more speed. |
|
|
32
|
+
| **ponytail-help** | `/ponytail-help` | This card. |
|
|
33
|
+
|
|
34
|
+
Codex uses `@ponytail`, `@ponytail-review`, and `@ponytail-help`; Claude Code
|
|
35
|
+
and OpenCode use the slash-command forms above (OpenCode ships `/ponytail` and
|
|
36
|
+
`/ponytail-review`).
|
|
37
|
+
|
|
38
|
+
## Deactivate
|
|
39
|
+
|
|
40
|
+
Say "stop ponytail" or "normal mode". Resume anytime with `/ponytail`.
|
|
41
|
+
`/ponytail off` also works.
|
|
42
|
+
|
|
43
|
+
## Configure Default Mode
|
|
44
|
+
|
|
45
|
+
Default mode = `full`, auto-active every session. Change it:
|
|
46
|
+
|
|
47
|
+
**Environment variable** (highest priority):
|
|
48
|
+
```bash
|
|
49
|
+
export PONYTAIL_DEFAULT_MODE=ultra
|
|
50
|
+
```
|
|
51
|
+
|
|
52
|
+
**Config file** (`~/.config/ponytail/config.json`, Windows: `%APPDATA%\ponytail\config.json`):
|
|
53
|
+
```json
|
|
54
|
+
{ "defaultMode": "lite" }
|
|
55
|
+
```
|
|
56
|
+
|
|
57
|
+
Set `"off"` to disable auto-activation on session start, activate manually
|
|
58
|
+
with `/ponytail` when wanted.
|
|
59
|
+
|
|
60
|
+
Resolution: env var > config file > `full`.
|
|
61
|
+
|
|
62
|
+
## Update
|
|
63
|
+
|
|
64
|
+
Enable auto-update once: open `/plugin`, go to Marketplaces, pick ponytail, Enable auto-update. Claude Code then pulls new versions at startup (run `/reload-plugins` when it prompts). Manual refresh: `/plugin marketplace update ponytail` then `/reload-plugins`.
|
|
65
|
+
|
|
66
|
+
If `/plugin` is not recognized, your Claude Code is out of date. Update it (`npm install -g @anthropic-ai/claude-code@latest`, or `brew upgrade claude-code`) and restart. Other hosts use their own update flow.
|
|
67
|
+
|
|
68
|
+
## More
|
|
69
|
+
|
|
70
|
+
Full docs + examples: https://github.com/DietrichGebert/ponytail
|
|
@@ -1,58 +1,58 @@
|
|
|
1
|
-
---
|
|
2
|
-
name: ponytail-review
|
|
3
|
-
description: >
|
|
4
|
-
Code review focused exclusively on over-engineering. Finds what to delete:
|
|
5
|
-
reinvented standard library, unneeded dependencies, speculative abstractions,
|
|
6
|
-
dead flexibility. One line per finding: location, what to cut, what replaces
|
|
7
|
-
it. Use when the user says "review for over-engineering", "what can we
|
|
8
|
-
delete", "is this over-engineered", "simplify review", or invokes
|
|
9
|
-
/ponytail-review. Complements correctness-focused review, this one only
|
|
10
|
-
hunts complexity.
|
|
11
|
-
disable-model-invocation: true
|
|
12
|
-
---
|
|
13
|
-
|
|
14
|
-
Review diffs for unnecessary complexity. One line per finding: location, what
|
|
15
|
-
to cut, what replaces it. The diff's best outcome is getting shorter.
|
|
16
|
-
|
|
17
|
-
## Format
|
|
18
|
-
|
|
19
|
-
`L<line>: <tag> <what>. <replacement>.`, or `<file>:L<line>: ...` for
|
|
20
|
-
multi-file diffs.
|
|
21
|
-
|
|
22
|
-
Tags:
|
|
23
|
-
|
|
24
|
-
- `delete:` dead code, unused flexibility, speculative feature. Replacement: nothing.
|
|
25
|
-
- `stdlib:` hand-rolled thing the standard library ships. Name the function.
|
|
26
|
-
- `native:` dependency or code doing what the platform already does. Name the feature.
|
|
27
|
-
- `yagni:` abstraction with one implementation, config nobody sets, layer with one caller.
|
|
28
|
-
- `shrink:` same logic, fewer lines. Show the shorter form.
|
|
29
|
-
|
|
30
|
-
## Examples
|
|
31
|
-
|
|
32
|
-
❌ "This EmailValidator class might be more complex than necessary, have you
|
|
33
|
-
considered whether all these validation rules are needed at this stage?"
|
|
34
|
-
|
|
35
|
-
✅ `L12-38: stdlib: 27-line validator class. "@" in email, 1 line, real validation is the confirmation mail.`
|
|
36
|
-
|
|
37
|
-
✅ `L4: native: moment.js imported for one format call. Intl.DateTimeFormat, 0 deps.`
|
|
38
|
-
|
|
39
|
-
✅ `repo.py:L88: yagni: AbstractRepository with one implementation. Inline it until a second one exists.`
|
|
40
|
-
|
|
41
|
-
✅ `L52-71: delete: retry wrapper around an idempotent local call. Nothing replaces it.`
|
|
42
|
-
|
|
43
|
-
✅ `L30-44: shrink: manual loop builds dict. dict(zip(keys, values)), 1 line.`
|
|
44
|
-
|
|
45
|
-
## Scoring
|
|
46
|
-
|
|
47
|
-
End with the only metric that matters: `net: -<N> lines possible.`
|
|
48
|
-
|
|
49
|
-
If there is nothing to cut, say `Lean already. Ship.` and stop.
|
|
50
|
-
|
|
51
|
-
## Boundaries
|
|
52
|
-
|
|
53
|
-
Scope: over-engineering and complexity only. Correctness bugs, security holes,
|
|
54
|
-
and performance are explicitly out of scope. Route them to a normal review
|
|
55
|
-
pass, not this one. A single smoke test or `assert`-based
|
|
56
|
-
self-check is the ponytail minimum, not bloat, never flag it for deletion.
|
|
57
|
-
Does not apply the fixes, only lists them.
|
|
58
|
-
"stop ponytail-review" or "normal mode": revert to verbose review style.
|
|
1
|
+
---
|
|
2
|
+
name: ponytail-review
|
|
3
|
+
description: >
|
|
4
|
+
Code review focused exclusively on over-engineering. Finds what to delete:
|
|
5
|
+
reinvented standard library, unneeded dependencies, speculative abstractions,
|
|
6
|
+
dead flexibility. One line per finding: location, what to cut, what replaces
|
|
7
|
+
it. Use when the user says "review for over-engineering", "what can we
|
|
8
|
+
delete", "is this over-engineered", "simplify review", or invokes
|
|
9
|
+
/ponytail-review. Complements correctness-focused review, this one only
|
|
10
|
+
hunts complexity.
|
|
11
|
+
disable-model-invocation: true
|
|
12
|
+
---
|
|
13
|
+
|
|
14
|
+
Review diffs for unnecessary complexity. One line per finding: location, what
|
|
15
|
+
to cut, what replaces it. The diff's best outcome is getting shorter.
|
|
16
|
+
|
|
17
|
+
## Format
|
|
18
|
+
|
|
19
|
+
`L<line>: <tag> <what>. <replacement>.`, or `<file>:L<line>: ...` for
|
|
20
|
+
multi-file diffs.
|
|
21
|
+
|
|
22
|
+
Tags:
|
|
23
|
+
|
|
24
|
+
- `delete:` dead code, unused flexibility, speculative feature. Replacement: nothing.
|
|
25
|
+
- `stdlib:` hand-rolled thing the standard library ships. Name the function.
|
|
26
|
+
- `native:` dependency or code doing what the platform already does. Name the feature.
|
|
27
|
+
- `yagni:` abstraction with one implementation, config nobody sets, layer with one caller.
|
|
28
|
+
- `shrink:` same logic, fewer lines. Show the shorter form.
|
|
29
|
+
|
|
30
|
+
## Examples
|
|
31
|
+
|
|
32
|
+
❌ "This EmailValidator class might be more complex than necessary, have you
|
|
33
|
+
considered whether all these validation rules are needed at this stage?"
|
|
34
|
+
|
|
35
|
+
✅ `L12-38: stdlib: 27-line validator class. "@" in email, 1 line, real validation is the confirmation mail.`
|
|
36
|
+
|
|
37
|
+
✅ `L4: native: moment.js imported for one format call. Intl.DateTimeFormat, 0 deps.`
|
|
38
|
+
|
|
39
|
+
✅ `repo.py:L88: yagni: AbstractRepository with one implementation. Inline it until a second one exists.`
|
|
40
|
+
|
|
41
|
+
✅ `L52-71: delete: retry wrapper around an idempotent local call. Nothing replaces it.`
|
|
42
|
+
|
|
43
|
+
✅ `L30-44: shrink: manual loop builds dict. dict(zip(keys, values)), 1 line.`
|
|
44
|
+
|
|
45
|
+
## Scoring
|
|
46
|
+
|
|
47
|
+
End with the only metric that matters: `net: -<N> lines possible.`
|
|
48
|
+
|
|
49
|
+
If there is nothing to cut, say `Lean already. Ship.` and stop.
|
|
50
|
+
|
|
51
|
+
## Boundaries
|
|
52
|
+
|
|
53
|
+
Scope: over-engineering and complexity only. Correctness bugs, security holes,
|
|
54
|
+
and performance are explicitly out of scope. Route them to a normal review
|
|
55
|
+
pass, not this one. A single smoke test or `assert`-based
|
|
56
|
+
self-check is the ponytail minimum, not bloat, never flag it for deletion.
|
|
57
|
+
Does not apply the fixes, only lists them.
|
|
58
|
+
"stop ponytail-review" or "normal mode": revert to verbose review style.
|
|
@@ -1,20 +1,20 @@
|
|
|
1
|
-
---
|
|
2
|
-
name: selesai-handoff
|
|
3
|
-
description: Hand the current conversation off to a fresh background agent that picks up the work immediately.
|
|
4
|
-
argument-hint: "What will the next session be used for?"
|
|
5
|
-
disable-model-invocation: true
|
|
6
|
-
---
|
|
7
|
-
|
|
8
|
-
Write a handoff summary of the current conversation so a fresh agent can continue the work. Instead of saving it, launch a background agent seeded with the summary as its prompt: `selesai --name "<descriptive name>" -p "<handoff summary>"`. It starts in the current working directory and returns immediately; the user manages it with `selesai agents`.
|
|
9
|
-
|
|
10
|
-
Always pass `-n`/`--name` with a descriptive name (e.g. `--name "Fix login bug"`) — it sets the display name shown in the job list, session picker, and terminal title.
|
|
11
|
-
|
|
12
|
-
You might want to use something like `nohup` or `Start-Process` or `start /b` or etc. before the selesai command to spawn background command.
|
|
13
|
-
|
|
14
|
-
Include a "suggested skills" section in the summary, which suggests skills that the agent should invoke.
|
|
15
|
-
|
|
16
|
-
Do not duplicate content already captured in other artifacts (PRDs, plans, ADRs, issues, commits, diffs). Reference them by path or URL instead.
|
|
17
|
-
|
|
18
|
-
Redact any sensitive information, such as API keys, passwords, or personally identifiable information — the summary becomes the agent's prompt.
|
|
19
|
-
|
|
20
|
-
If the user passed arguments, treat them as a description of what the next session will focus on and tailor the summary accordingly.
|
|
1
|
+
---
|
|
2
|
+
name: selesai-handoff
|
|
3
|
+
description: Hand the current conversation off to a fresh background agent that picks up the work immediately.
|
|
4
|
+
argument-hint: "What will the next session be used for?"
|
|
5
|
+
disable-model-invocation: true
|
|
6
|
+
---
|
|
7
|
+
|
|
8
|
+
Write a handoff summary of the current conversation so a fresh agent can continue the work. Instead of saving it, launch a background agent seeded with the summary as its prompt: `selesai --name "<descriptive name>" -p "<handoff summary>"`. It starts in the current working directory and returns immediately; the user manages it with `selesai agents`.
|
|
9
|
+
|
|
10
|
+
Always pass `-n`/`--name` with a descriptive name (e.g. `--name "Fix login bug"`) — it sets the display name shown in the job list, session picker, and terminal title.
|
|
11
|
+
|
|
12
|
+
You might want to use something like `nohup` or `Start-Process` or `start /b` or etc. before the selesai command to spawn background command.
|
|
13
|
+
|
|
14
|
+
Include a "suggested skills" section in the summary, which suggests skills that the agent should invoke.
|
|
15
|
+
|
|
16
|
+
Do not duplicate content already captured in other artifacts (PRDs, plans, ADRs, issues, commits, diffs). Reference them by path or URL instead.
|
|
17
|
+
|
|
18
|
+
Redact any sensitive information, such as API keys, passwords, or personally identifiable information — the summary becomes the agent's prompt.
|
|
19
|
+
|
|
20
|
+
If the user passed arguments, treat them as a description of what the next session will focus on and tailor the summary accordingly.
|
|
@@ -1,73 +1,73 @@
|
|
|
1
|
-
---
|
|
2
|
-
name: workflow-creation
|
|
3
|
-
description: Creates durable Selesai workflow modes. Use when a user asks to create or change a workflow mode, phased agent flow, or slash-command workflow.
|
|
4
|
-
disable-model-invocation: true
|
|
5
|
-
---
|
|
6
|
-
|
|
7
|
-
# Durable Workflows
|
|
8
|
-
|
|
9
|
-
**Durable** means `workflow.json`, not session history, is the run authority. Build on the shared engine in `src/extensions/workflow/`; a mode is configuration, not a second orchestrator.
|
|
10
|
-
|
|
11
|
-
## 1. Choose the smallest fit
|
|
12
|
-
|
|
13
|
-
Read `docs/workflows.md`, every file in `src/extensions/workflow/modes/`, and the relevant workflow tests.
|
|
14
|
-
|
|
15
|
-
- Reuse `prototype`, `quick`, or `task` when its phase graph and terminal gate fit; change only its prompts/configuration.
|
|
16
|
-
- Add a mode only for a materially different graph, artifact contract, or close gate.
|
|
17
|
-
|
|
18
|
-
Done when the request is mapped to one existing mode or a named new mode with its phase list and terminal artifact.
|
|
19
|
-
|
|
20
|
-
## 2. Trace the durable seam
|
|
21
|
-
|
|
22
|
-
Before changing engine-facing behavior, read completely:
|
|
23
|
-
|
|
24
|
-
- `state-machine.ts` — graph, artifact gates, terminal-ready, completion;
|
|
25
|
-
- `adapter.ts` — tools, event handlers, persistence, reload guards;
|
|
26
|
-
- `run-state.ts` — canonical record and resume validation;
|
|
27
|
-
- `extension.ts` — single extension mounting all modes.
|
|
28
|
-
|
|
29
|
-
Done when every proposed behavior has one owner: state machine, shared adapter, or mode configuration.
|
|
30
|
-
|
|
31
|
-
## 3. Implement the mode
|
|
32
|
-
|
|
33
|
-
For a new mode, add `src/extensions/workflow/modes/<name>.ts`, modeled on `quick.ts`, with only `WorkflowConfig` and `WorkflowModeRegistration`:
|
|
34
|
-
|
|
35
|
-
- ordered phases, phase artifacts, prompts, validators, close artifacts/validators;
|
|
36
|
-
- unique mode/status/entry identities and slash-command name; users start/resume through that command, while the shared end tool selects the mode;
|
|
37
|
-
- prompts that name exact artifact paths and use `write_workflow_artifact` only for workflow artifacts.
|
|
38
|
-
|
|
39
|
-
Register the mode once in `MODES` in `extension.ts`; document its lifecycle and commands in `docs/workflows.md`.
|
|
40
|
-
|
|
41
|
-
Done when the mode file has no filesystem, persistence, event-registration, or controller code.
|
|
42
|
-
|
|
43
|
-
## 4. Preserve the durable contract
|
|
44
|
-
|
|
45
|
-
The shared adapter owns UUID artifact directories, atomic saves, resume, loop review state, and reload safety. Do not reimplement them per mode.
|
|
46
|
-
|
|
47
|
-
- Persisted state changes after start, artifact/loop transition, resume reconciliation, and explicit end.
|
|
48
|
-
- Never auto-resume on `session_start`; only a user-invoked mode command with an explicit selector attaches a run.
|
|
49
|
-
- Artifact completion advances durable state then stops the parent turn; the user deliberately continues the attached mode.
|
|
50
|
-
- Terminal-ready stays active. Only `end_workflow({ mode })` marks the record completed and terminates.
|
|
51
|
-
- One `ExtensionAPI` hosts all modes: shared writer once, stale reload handlers inert, one attached run total.
|
|
52
|
-
- Builder/reviewer loops use adapter-owned rounds, review files, markers, and max-iteration pause.
|
|
53
|
-
|
|
54
|
-
Done when the new behavior preserves every applicable invariant above.
|
|
55
|
-
|
|
56
|
-
## 5. Lock the mode with real seams
|
|
57
|
-
|
|
58
|
-
Extend the existing fake-Pi tests; do not add another framework. Cover the real mode, not only state-machine units:
|
|
59
|
-
|
|
60
|
-
- start writes valid `workflow.json`; artifact transition updates it; explicit end completes it;
|
|
61
|
-
- explicit resume reconciles an artifact written before a phase save;
|
|
62
|
-
- loop round/review path resumes when the mode has a loop;
|
|
63
|
-
- terminal-ready does not complete early;
|
|
64
|
-
- extension reload ignores stale handlers; inactive sibling modes do not react.
|
|
65
|
-
|
|
66
|
-
Run the narrow mode test, then:
|
|
67
|
-
|
|
68
|
-
```bash
|
|
69
|
-
npx vitest run src/__tests__/state-machine.test.ts src/__tests__/adapter.test.ts src/__tests__/workflow-race.test.ts src/__tests__/workflow-run-state.test.ts src/__tests__/<mode>-workflow.test.ts
|
|
70
|
-
npm run build
|
|
71
|
-
```
|
|
72
|
-
|
|
73
|
-
Done when those checks pass and the diff contains only the selected mode, shared-engine changes proven necessary by a failing regression, documentation, and tests.
|
|
1
|
+
---
|
|
2
|
+
name: workflow-creation
|
|
3
|
+
description: Creates durable Selesai workflow modes. Use when a user asks to create or change a workflow mode, phased agent flow, or slash-command workflow.
|
|
4
|
+
disable-model-invocation: true
|
|
5
|
+
---
|
|
6
|
+
|
|
7
|
+
# Durable Workflows
|
|
8
|
+
|
|
9
|
+
**Durable** means `workflow.json`, not session history, is the run authority. Build on the shared engine in `src/extensions/workflow/`; a mode is configuration, not a second orchestrator.
|
|
10
|
+
|
|
11
|
+
## 1. Choose the smallest fit
|
|
12
|
+
|
|
13
|
+
Read `docs/workflows.md`, every file in `src/extensions/workflow/modes/`, and the relevant workflow tests.
|
|
14
|
+
|
|
15
|
+
- Reuse `prototype`, `quick`, or `task` when its phase graph and terminal gate fit; change only its prompts/configuration.
|
|
16
|
+
- Add a mode only for a materially different graph, artifact contract, or close gate.
|
|
17
|
+
|
|
18
|
+
Done when the request is mapped to one existing mode or a named new mode with its phase list and terminal artifact.
|
|
19
|
+
|
|
20
|
+
## 2. Trace the durable seam
|
|
21
|
+
|
|
22
|
+
Before changing engine-facing behavior, read completely:
|
|
23
|
+
|
|
24
|
+
- `state-machine.ts` — graph, artifact gates, terminal-ready, completion;
|
|
25
|
+
- `adapter.ts` — tools, event handlers, persistence, reload guards;
|
|
26
|
+
- `run-state.ts` — canonical record and resume validation;
|
|
27
|
+
- `extension.ts` — single extension mounting all modes.
|
|
28
|
+
|
|
29
|
+
Done when every proposed behavior has one owner: state machine, shared adapter, or mode configuration.
|
|
30
|
+
|
|
31
|
+
## 3. Implement the mode
|
|
32
|
+
|
|
33
|
+
For a new mode, add `src/extensions/workflow/modes/<name>.ts`, modeled on `quick.ts`, with only `WorkflowConfig` and `WorkflowModeRegistration`:
|
|
34
|
+
|
|
35
|
+
- ordered phases, phase artifacts, prompts, validators, close artifacts/validators;
|
|
36
|
+
- unique mode/status/entry identities and slash-command name; users start/resume through that command, while the shared end tool selects the mode;
|
|
37
|
+
- prompts that name exact artifact paths and use `write_workflow_artifact` only for workflow artifacts.
|
|
38
|
+
|
|
39
|
+
Register the mode once in `MODES` in `extension.ts`; document its lifecycle and commands in `docs/workflows.md`.
|
|
40
|
+
|
|
41
|
+
Done when the mode file has no filesystem, persistence, event-registration, or controller code.
|
|
42
|
+
|
|
43
|
+
## 4. Preserve the durable contract
|
|
44
|
+
|
|
45
|
+
The shared adapter owns UUID artifact directories, atomic saves, resume, loop review state, and reload safety. Do not reimplement them per mode.
|
|
46
|
+
|
|
47
|
+
- Persisted state changes after start, artifact/loop transition, resume reconciliation, and explicit end.
|
|
48
|
+
- Never auto-resume on `session_start`; only a user-invoked mode command with an explicit selector attaches a run.
|
|
49
|
+
- Artifact completion advances durable state then stops the parent turn; the user deliberately continues the attached mode.
|
|
50
|
+
- Terminal-ready stays active. Only `end_workflow({ mode })` marks the record completed and terminates.
|
|
51
|
+
- One `ExtensionAPI` hosts all modes: shared writer once, stale reload handlers inert, one attached run total.
|
|
52
|
+
- Builder/reviewer loops use adapter-owned rounds, review files, markers, and max-iteration pause.
|
|
53
|
+
|
|
54
|
+
Done when the new behavior preserves every applicable invariant above.
|
|
55
|
+
|
|
56
|
+
## 5. Lock the mode with real seams
|
|
57
|
+
|
|
58
|
+
Extend the existing fake-Pi tests; do not add another framework. Cover the real mode, not only state-machine units:
|
|
59
|
+
|
|
60
|
+
- start writes valid `workflow.json`; artifact transition updates it; explicit end completes it;
|
|
61
|
+
- explicit resume reconciles an artifact written before a phase save;
|
|
62
|
+
- loop round/review path resumes when the mode has a loop;
|
|
63
|
+
- terminal-ready does not complete early;
|
|
64
|
+
- extension reload ignores stale handlers; inactive sibling modes do not react.
|
|
65
|
+
|
|
66
|
+
Run the narrow mode test, then:
|
|
67
|
+
|
|
68
|
+
```bash
|
|
69
|
+
npx vitest run src/__tests__/state-machine.test.ts src/__tests__/adapter.test.ts src/__tests__/workflow-race.test.ts src/__tests__/workflow-run-state.test.ts src/__tests__/<mode>-workflow.test.ts
|
|
70
|
+
npm run build
|
|
71
|
+
```
|
|
72
|
+
|
|
73
|
+
Done when those checks pass and the diff contains only the selected mode, shared-engine changes proven necessary by a failing regression, documentation, and tests.
|