@kolbo/mcp 1.90.0 → 1.90.1
This diff represents the content of publicly available package versions that have been released to one of the supported registries. The information contained in this diff is provided for informational purposes only and reflects changes between package versions as they appear in their respective public registries.
- package/package.json +1 -1
- package/skill/GENERATED.md +1 -1
- package/skill/SKILL.md +1 -1
- package/skill/references/models/html-presentation.md +1 -1
- package/skill/references/models/landing-page.md +1 -1
- package/skill/references/models/seedance25.md +1 -1
- package/skill/references/models/visual-code.md +1 -1
- package/skill/references/workflows/cost-and-validation.md +4 -7
- package/skill/references/workflows/filmmaking.md +1 -1
- package/skill/references/workflows/troubleshooting.md +2 -2
- package/skill/references/workflows/video-editor.md +1 -1
package/package.json
CHANGED
package/skill/GENERATED.md
CHANGED
|
@@ -1,6 +1,6 @@
|
|
|
1
1
|
# AUTO-GENERATED — do not edit
|
|
2
2
|
|
|
3
|
-
This tree is mirrored from kolbo-code@
|
|
3
|
+
This tree is mirrored from kolbo-code@b77c8b4, the single source of truth.
|
|
4
4
|
Canonical source: packages/opencode/skills/kolbo/
|
|
5
5
|
Distribution: .github/workflows/sync-skill-to-plugin.yml
|
|
6
6
|
|
package/skill/SKILL.md
CHANGED
|
@@ -445,7 +445,7 @@ HTML/SVG/Mermaid artifacts have a **Share** button in the preview toolbar that u
|
|
|
445
445
|
If at this point you still don't know which `references/` file to load, default to `references/models/prompt-copilot.md` for generation prompts or `references/workflows/cost-and-validation.md` for cost/validation questions, or just keep going with this core file's rules.
|
|
446
446
|
|
|
447
447
|
## Media selection preferences
|
|
448
|
-
Honor explicit models, presets, budget and inputs. Choose only eligible catalog candidates with all required capabilities. Use requested presets; otherwise use fitting presets when useful. For
|
|
448
|
+
Honor explicit models, presets, budget and inputs. Choose only eligible catalog candidates with all required capabilities. Use requested presets; otherwise use fitting presets when useful. For video generation, editing and lip-sync, when the user has not explicitly selected an output resolution, use the cheapest supported output resolution from the live catalog and pass it explicitly; do not inherit an expensive provider default. Preserve explicit user-selected resolution/settings. Finish fully, cinematic, professional, final, production and available credits are NOT permission to increase resolution. Never infer output resolution from reference media or export settings. A budget is a ceiling, not a spending target. Do not upscale or regenerate at a higher tier without explicit user authorization. If pricing or supported resolutions cannot be verified, inspect the catalog before dispatch; never invent a tier. Models with fixed output resolution use their native output. Never treat a policy refusal as a technical failure or route around safeguards.
|
|
449
449
|
Default images and edits: GPT Image 2.5 Flare/Sunburst; medium for value, high for ordinary maximum quality. Reserve xhigh/max for exceptional dense or difficult multilingual text after medium/high prove insufficient; do not automatically spend on retries. Nano Banana 2 is secondary. Seedream 5.0 Pro favors cinematic aesthetics over complex instruction fidelity; Wan 2.7 Pro is another creative alternative. Z Image/P Image for cheap tests. Midjourney for artistic concepts only, never editing. Soul V2 for realistic people/UGC concepts; derive character sheets before registering finished Visual DNA. Mirage Film 2 for environments and cinematic inspiration.
|
|
450
450
|
Default video: Seedance 2.5. Kling specializes in controlled single-image and first/last-frame shots. Wan 3.0 specializes in motion graphics and Hebrew/dialogue work. MiniMax H3 offers higher resolution; H3 Max favors speed at lower resolution. Gemini Omni Flash is a secondary Hebrew option (up to 10 seconds); Grok Imagine 1.5 and Seedance 2.0 are alternatives. P Video/Draft for cheap fast tests. Use base, edit or extend variants only with their required inputs.
|
|
451
451
|
Existing-video lip-sync: Sync 3 for active-speaker handling; PixVerse for cartoons/2D and economical faster work. Portrait lip-sync: Veed Fabric or HeyGen Avatar; P Avatar for budget work. LTX Audio to Video for camera/environment motion with audio-driven performance.
|
|
@@ -6,7 +6,7 @@
|
|
|
6
6
|
|
|
7
7
|
Load this file when the user wants to **build / create / generate an HTML presentation, slide deck, or pitch deck**. For landing pages see `models/landing-page.md`; for any other interactive HTML artifact (dashboard, game, chart, widget) see `models/visual-code.md`.
|
|
8
8
|
|
|
9
|
-
**
|
|
9
|
+
**Kobi Code routing:** write the artifact as a single HTML block in your reply. The Kobi Code panel renders it as a previewable artifact card. Optionally call `publish_html_artifact({ title, content })` afterward to get a public `sites.kolbo.ai` URL.
|
|
10
10
|
|
|
11
11
|
## 🚨 NON-NEGOTIABLE: Viewport Fitting
|
|
12
12
|
|
|
@@ -6,7 +6,7 @@
|
|
|
6
6
|
|
|
7
7
|
Load this file when the user wants to **build / create a landing page, marketing site, one-pager, product page, app launch page, SaaS sign-up page, or event page**. For slide decks see `models/html-presentation.md`; for dashboards / games / charts / widgets see `models/visual-code.md`.
|
|
8
8
|
|
|
9
|
-
**
|
|
9
|
+
**Kobi Code routing:** write the artifact as a single HTML block in your reply. Kobi Code's panel renders it as a previewable artifact card. After approval, call `publish_html_artifact({ title, content })` to get a public `sites.kolbo.ai` URL.
|
|
10
10
|
|
|
11
11
|
## 🎯 Design Thinking — Commit Before You Code
|
|
12
12
|
|
|
@@ -13,7 +13,7 @@ Load this file when the user wants a **Seedance 2.5** video (they said "2.5" / "
|
|
|
13
13
|
|
|
14
14
|
**Dialogue language: English.** Other languages are not reliably performed, and Hebrew does not work — it returns accented gibberish or English-shaped mouth movement. Never offer a user "Hebrew dialogue directly". This restriction applies to rendered dialogue/prose only, never to binding identifiers: preserve an exact stored Hebrew Visual DNA or moodboard tag such as `@אביב` / `#ישראל` literally. See `models/seedance.md` for the three honest alternatives.
|
|
15
15
|
|
|
16
|
-
**
|
|
16
|
+
**Use the cheapest supported tier unless the user selected an output resolution.** Resolution is a credit MULTIPLIER, not a flat rate. Relative to 720p: 480p ×0.44, 1080p ×2.25. A 30s pass costs ~540cr at 480p against ~1230cr at 720p and ~2770cr at 1080p. When no output resolution was selected and 480p is the cheapest supported tier, block the film at 480p, get the user's sign-off on staging, performance and timing, then re-run only the approved cut at a higher delivery resolution if the user explicitly authorizes that resolution increase. Approval of the creative cut alone does not authorize a more expensive resolution. If no output resolution was selected, use the cheapest supported tier from the live catalog even for final work; pass it explicitly.
|
|
17
17
|
|
|
18
18
|
## What's NEW in 2.5 (verified — never hedge)
|
|
19
19
|
|
|
@@ -8,7 +8,7 @@ Load this file when the user wants to **build an interactive HTML artifact where
|
|
|
8
8
|
|
|
9
9
|
If the user asks for a **presentation** → see `models/html-presentation.md`. If they ask for a **landing page** → see `models/landing-page.md`. Everything else visual-and-interactive is here.
|
|
10
10
|
|
|
11
|
-
**
|
|
11
|
+
**Kobi Code routing:** write the artifact as a single HTML block in your reply. Kobi Code's panel renders it as a previewable artifact card. Call `publish_html_artifact({ title, content })` to publish to `sites.kolbo.ai` after approval.
|
|
12
12
|
|
|
13
13
|
## What This Skill Is For
|
|
14
14
|
|
|
@@ -102,18 +102,15 @@ Normal cost formula: `final_cost = credit × output_seconds × resolution_multip
|
|
|
102
102
|
- ✅ Show them the supported set in one line and ask:
|
|
103
103
|
> "Seedance 2 elements supports `[720p, 1080p, 1440p, 2160p]` — 480p isn't available. Closest cheap option is 720p (~+0 credits over your intent). Want 720p, or pick another?"
|
|
104
104
|
- Only fire after they reply.
|
|
105
|
-
2. **
|
|
106
|
-
|
|
107
|
-
|
|
108
|
-
- final / production / hero → highest the user's budget allows (3K-4K / 1440p-2160p)
|
|
109
|
-
3. **No quality signal AND cost difference >2×** OR total batch ≥4 outputs → **ask the user once** with a one-line cost comparison, then default to standard if they don't reply.
|
|
110
|
-
4. **No quality signal AND cost difference ≤1.5×** → quietly use the cheapest supported, no need to interrupt.
|
|
105
|
+
2. **No explicit video output resolution**: choose the cheapest supported tier using current catalog pricing and pass it explicitly. This applies to drafts, normal work and final delivery alike. Do not default to 720p/1080p when a cheaper supported tier exists. Fixed-resolution models use their native output.
|
|
106
|
+
3. **Creative intent is not spending authorization**: "finish fully", "cinematic", "professional", "final", "production", "hero" and "don't ask me" do not authorize higher resolution, upscaling or a second high-resolution generation. A budget is a ceiling, not a target. Reference-video resolution and export resolution do not authorize matching generation resolution.
|
|
107
|
+
4. Preserve explicit user-selected settings. Otherwise proceed economically without a resolution approval loop. Inspect missing pricing/capabilities before dispatch. Upgrade only when the user explicitly selects a higher output tier or authorizes the resolution increase; never treat silence as approval. Image quality follows the image-model guidance (GPT Image 2.5 medium by default), not a generic final-work maximum.
|
|
111
108
|
5. **Sound on a video model with `sound_credit_multiplier > 1`** → if user didn't ask for sound, leave it off. If user said "with sound" / "with music", enable it.
|
|
112
109
|
|
|
113
110
|
## Defaults When Nothing Is Specified
|
|
114
111
|
|
|
115
112
|
- **Image**: `1K` (or the cheapest in `supported_resolutions`).
|
|
116
|
-
- **Video**:
|
|
113
|
+
- **Video**: cheapest supported output resolution from current catalog pricing, passed explicitly. Use the duration required by the user/task; do not lengthen clips to spend the available budget.
|
|
117
114
|
- **Sound**: respect `sound_enabled_by_default`; if false, leave off.
|
|
118
115
|
|
|
119
116
|
## Log Approved Resolution / Duration / Sound Choices
|
|
@@ -145,4 +145,4 @@ Fix errors before delivery. Report warnings that represent genuine creative trad
|
|
|
145
145
|
|
|
146
146
|
Read [workflows.md](references/filmmaking/workflows.md) for single shots, dialogue scenes, music performance, connected sequences, impossible shots, and feature workflows.
|
|
147
147
|
|
|
148
|
-
This workflow is part of the canonical Kolbo skill. The
|
|
148
|
+
This workflow is part of the canonical Kolbo skill. The Kobi Code sync pipeline mirrors it to MCP and plugin consumers; product surfaces may compile the same filmmaking truth through their own model adapters.
|
|
@@ -73,9 +73,9 @@ Branch on `failure.category` / `failure.retryable`:
|
|
|
73
73
|
- `retryable === true` (transient: network, rate limit, provider 5xx) → retry once with the same payload after a short pause. If it fails again, surface to user.
|
|
74
74
|
- `retryable === false` and unknown category → surface the raw `message` to the user, don't retry.
|
|
75
75
|
|
|
76
|
-
##
|
|
76
|
+
## Kobi Code Documentation
|
|
77
77
|
|
|
78
|
-
Full public documentation for
|
|
78
|
+
Full public documentation for Kobi Code (the CLI you are running inside) lives at **[docs.kolbo.ai/docs/kolbo-code](https://docs.kolbo.ai/docs/kolbo-code)**. If the user asks about installation, authentication, voice input, supported languages, commands, or how to uninstall, point them to the matching page below rather than guessing:
|
|
79
79
|
|
|
80
80
|
| Topic | Path |
|
|
81
81
|
|-------|------|
|
|
@@ -23,6 +23,6 @@ For creation, `create_video_editor_session` accepts the original clips/audio/tex
|
|
|
23
23
|
- Caption items support content, words, typography, RTL, background, stroke, shadow and captionStyle. Word startTime/endTime use absolute timeline milliseconds and must fit inside the caption item. Moving/reordering captions moves their words too. To derive captions from speech, use `transcribe_audio` first; edits themselves do not transcribe or generate media.
|
|
24
24
|
- Settings include transforms, opacity, visibility, audio volume/gain/fades, playback rate, grading, motion and the available text/caption effects. Consult the schema rather than inventing fields.
|
|
25
25
|
|
|
26
|
-
The tools edit saved state.
|
|
26
|
+
The tools edit saved state. Updated browser editors receive each saved change automatically and preserve the playhead and view. Unsaved manual edits pause autosave and show a conflict choice. Browser and agent writes both use revision checks. This requires the updated API and browser; it does not stream an agent's unsaved intermediate operations.
|
|
27
27
|
|
|
28
28
|
Saved settings and successful exports do not prove visual parity for every renderer effect. Inspect the output before claiming visual quality. Respect the user's deployment, publishing and generation-spend boundaries.
|