makaron-cli 0.14.8 → 0.14.10
This diff represents the content of publicly available package versions that have been released to one of the supported registries. The information contained in this diff is provided for informational purposes only and reflects changes between package versions as they appear in their respective public registries.
- package/.claude-plugin/plugin.json +1 -1
- package/.codex-plugin/plugin.json +1 -1
- package/README.md +6 -4
- package/bin/makaron.mjs +6 -5
- package/package.json +1 -1
- package/skills/makaron/SKILL.md +6 -4
package/README.md
CHANGED
|
@@ -98,7 +98,7 @@ npx makaron-cli chat --project auto --image photo.jpg --json -b "make it cinemat
|
|
|
98
98
|
npx makaron-cli chat --project auto --image img1.jpg --image img2.jpg --json -b "combine these"
|
|
99
99
|
```
|
|
100
100
|
|
|
101
|
-
`chat` routes image and video models automatically, but you may select the Agent LLM with `--agent-model`. Accepted values are `auto`, the base model IDs (`gpt-5.6-terra`, `gpt-5.6-sol`, `gpt-5.6-luna`, `grok-4.6`, `deepseek-v4-pro`), and the personal-plan routes (`gpt-5.6-terra-codex-subscription`, `gpt-5.6-sol-codex-subscription`, `gpt-5.6-luna-codex-subscription`, `grok-4.6-grok-subscription`). For the configured owner, `auto` resolves to GPT-5.6 Terra through the personal Codex plan. Base GPT-5.6 IDs select Azure API and base `grok-4.6` selects OpenRouter API; the suffixed IDs select the corresponding personal plan explicitly. This flag changes only the reasoning/tool-calling Agent LLM.
|
|
101
|
+
`chat` routes image and video models automatically, but you may select the Agent LLM with `--agent-model`. Accepted values are `auto`, the base model IDs (`gpt-5.6-terra`, `gpt-5.6-sol`, `gpt-5.6-luna`, `grok-4.6`, `deepseek-v4-pro`, `deepseek-flash`), and the personal-plan routes (`gpt-5.6-terra-codex-subscription`, `gpt-5.6-sol-codex-subscription`, `gpt-5.6-luna-codex-subscription`, `grok-4.6-grok-subscription`). For the configured owner, `auto` resolves to GPT-5.6 Terra through the personal Codex plan. Base GPT-5.6 IDs select Azure API and base `grok-4.6` selects OpenRouter API; the suffixed IDs select the corresponding personal plan explicitly. This flag changes only the reasoning/tool-calling Agent LLM.
|
|
102
102
|
|
|
103
103
|
```bash
|
|
104
104
|
# Explicit lower-cost Agent LLM for a controlled comparison
|
|
@@ -363,11 +363,11 @@ npx makaron-cli edit --image photo.jpg --ref style.jpg "match this style"
|
|
|
363
363
|
# Output to file
|
|
364
364
|
npx makaron-cli edit --image photo.jpg --out result.jpg "make it dramatic"
|
|
365
365
|
|
|
366
|
-
# Strict transparent PNG/WebP output through GPT Image 2 (fails rather than returning opaque)
|
|
367
|
-
npx makaron-cli edit --image-model
|
|
366
|
+
# Strict transparent PNG/WebP output through GPT Image 2.5 Flare (fails rather than returning opaque)
|
|
367
|
+
npx makaron-cli edit --image-model gpt-image-2.5-flare --background transparent --out sticker.png "a magenta star sticker"
|
|
368
368
|
```
|
|
369
369
|
|
|
370
|
-
Options: `--image`, `--image-model gemini|gemini-lite|qwen|openai|wan2.7-image|pony|wai`, `--ref <file>` (up to 3), `--aspect <ratio>`, `--background auto|opaque|transparent`, `--out <path>`. Transparent output routes strictly to GPT Image 2 and is returned only when the provider supplies real PNG/WebP alpha.
|
|
370
|
+
Options: `--image`, `--image-model gemini|gemini-lite|qwen|openai|gpt-image-2.5-flare|gpt-image-2.5-sunburst|wan2.7-image|pony|wai`, `--ref <file>` (up to 3), `--aspect <ratio>`, `--background auto|opaque|transparent`, `--out <path>`. Transparent output routes strictly to GPT Image 2.5 Flare and is returned only when the provider supplies real PNG/WebP alpha.
|
|
371
371
|
|
|
372
372
|
`wan2.7-image` uses Alibaba international for fast, approximately 1K generation and editing (default 6 credits/image). Failed or timed-out Wan requests are not automatically retried or switched to another model. Face identity can change. Example: `makaron edit --image portrait.jpg --image-model wan2.7-image --aspect 16:9 --out stadium.jpg "Place this woman in a baseball stadium, preserving her face."`
|
|
373
373
|
|
|
@@ -656,3 +656,5 @@ npx makaron-cli admin fetch-skill https://www.makaron.app/s/4c4cbd57
|
|
|
656
656
|
- Before photo should match the person in the cover (hair, clothing, accessories)
|
|
657
657
|
|
|
658
658
|
FAL video models: **fal H3 Turbo** uses `minimax-h3-max` for single-start-frame I2V/T2V. **FAL H3 Max** uses the new selector `fal-h3-max`: native T2V or image/video/audio reference-to-video, default 768p, optional 480p/1080p, integer 5–15s; at most 9 images / 3 videos / 3 audios / 12 total. Reference video/audio each 2–15s and each modality totals at most 15s. Source-video modifications use generation with feature references, not typed edit/extend. Reference input tokens are billed in addition to output video; query current pricing.
|
|
659
|
+
|
|
660
|
+
The legacy `openai` image-model parameter now resolves to GPT Image 2.5 Flare.
|
package/bin/makaron.mjs
CHANGED
|
@@ -39,6 +39,7 @@ const CHAT_AGENT_MODELS = [
|
|
|
39
39
|
'grok-4.6',
|
|
40
40
|
'grok-4.6-grok-subscription',
|
|
41
41
|
'deepseek-v4-pro',
|
|
42
|
+
'deepseek-flash',
|
|
42
43
|
];
|
|
43
44
|
|
|
44
45
|
// Public anon key (safe to embed — only enables auth, not data access)
|
|
@@ -462,7 +463,7 @@ Options:
|
|
|
462
463
|
--media-manifest <file|-> Import typed image/video media before this run.
|
|
463
464
|
--skill <id|label|name> Use an installed skill or auto-install a matched marketplace skill.
|
|
464
465
|
--agent-model <id> Agent LLM only: auto, gpt-5.6-terra, gpt-5.6-sol,
|
|
465
|
-
gpt-5.6-luna, grok-4.6, deepseek-v4-pro, or a
|
|
466
|
+
gpt-5.6-luna, grok-4.6, deepseek-v4-pro, deepseek-flash, or a
|
|
466
467
|
gpt-5.6-*-codex-subscription or
|
|
467
468
|
grok-4.6-grok-subscription personal-plan route.
|
|
468
469
|
--background, -b Submit and print a runId.
|
|
@@ -1879,7 +1880,7 @@ Usage:
|
|
|
1879
1880
|
Options:
|
|
1880
1881
|
--image <file|url> Base image to edit. Omit for text-to-image.
|
|
1881
1882
|
--ref <file|url> Additional reference image. Repeatable, up to 3.
|
|
1882
|
-
--image-model <id> gemini, gemini-lite, qwen, openai, wan2.7-image, pony, or wai.
|
|
1883
|
+
--image-model <id> gemini, gemini-lite, qwen, openai, gpt-image-2.5-flare, gpt-image-2.5-sunburst, wan2.7-image, pony, or wai.
|
|
1883
1884
|
--skill <id> enhance, creative, wild, or captions.
|
|
1884
1885
|
--aspect <ratio> Output aspect ratio, for example 1:1, 16:9, or 9:16.
|
|
1885
1886
|
--background <mode> auto, opaque, or transparent.
|
|
@@ -1887,13 +1888,13 @@ Options:
|
|
|
1887
1888
|
--help, -h Show this help.
|
|
1888
1889
|
|
|
1889
1890
|
Notes:
|
|
1890
|
-
Model selection is optional. Transparent output routes strictly to GPT Image 2
|
|
1891
|
+
Model selection is optional. Transparent output routes strictly to GPT Image 2.5 Flare
|
|
1891
1892
|
and fails instead of returning an opaque fallback.
|
|
1892
1893
|
|
|
1893
1894
|
Examples:
|
|
1894
1895
|
makaron edit --image portrait.jpg --image-model qwen --out result.jpg "cinematic warm light"
|
|
1895
1896
|
makaron edit --image product.jpg --ref style.png --aspect 1:1 "use this visual style"
|
|
1896
|
-
makaron edit --image-model
|
|
1897
|
+
makaron edit --image-model gpt-image-2.5-flare --background transparent --out sticker.png "a magenta star sticker"
|
|
1897
1898
|
`);
|
|
1898
1899
|
}
|
|
1899
1900
|
|
|
@@ -2930,7 +2931,7 @@ if (!command || command === '--help' || command === '-h' || command === 'help')
|
|
|
2930
2931
|
else promptParts.push(args[i]);
|
|
2931
2932
|
}
|
|
2932
2933
|
editArgs.editPrompt = promptParts.join(' ');
|
|
2933
|
-
if (!editArgs.editPrompt) { console.error('Usage: makaron edit [--image <file|url>] [--image-model gemini|gemini-lite|qwen|openai|wan2.7-image|pony|wai] [--ref <file>] [--aspect <ratio>] [--background auto|opaque|transparent] [--out <file>] "prompt"'); process.exit(1); }
|
|
2934
|
+
if (!editArgs.editPrompt) { console.error('Usage: makaron edit [--image <file|url>] [--image-model gemini|gemini-lite|qwen|openai|gpt-image-2.5-flare|gpt-image-2.5-sunburst|wan2.7-image|pony|wai] [--ref <file>] [--aspect <ratio>] [--background auto|opaque|transparent] [--out <file>] "prompt"'); process.exit(1); }
|
|
2934
2935
|
process.stderr.write('🎨 Generating...\n');
|
|
2935
2936
|
const result = await callMcpTool(baseUrl, headers, 'makaron_edit_image', editArgs);
|
|
2936
2937
|
saveMcpImage(result, outputPath);
|
package/package.json
CHANGED
package/skills/makaron/SKILL.md
CHANGED
|
@@ -90,7 +90,7 @@ npx makaron-cli chat --project auto --image photo.jpg --json -b "make it cinemat
|
|
|
90
90
|
npx makaron-cli chat --project auto --image img1.jpg --image img2.jpg --json -b "combine these"
|
|
91
91
|
```
|
|
92
92
|
|
|
93
|
-
`chat` routes image and video models automatically. Use `--agent-model` only when the user explicitly asks to select or compare the reasoning/tool-calling Agent LLM. Accepted values are `auto`, the base model IDs (`gpt-5.6-terra`, `gpt-5.6-sol`, `gpt-5.6-luna`, `grok-4.6`, `deepseek-v4-pro`), and the personal-plan routes (`gpt-5.6-terra-codex-subscription`, `gpt-5.6-sol-codex-subscription`, `gpt-5.6-luna-codex-subscription`). For the configured owner, `auto` uses GPT-5.6 Terra through the personal Codex plan; base GPT-5.6 IDs select Azure API, while suffixed IDs explicitly select the personal plan. Never put an image or video model ID in `--agent-model`.
|
|
93
|
+
`chat` routes image and video models automatically. Use `--agent-model` only when the user explicitly asks to select or compare the reasoning/tool-calling Agent LLM. Accepted values are `auto`, the base model IDs (`gpt-5.6-terra`, `gpt-5.6-sol`, `gpt-5.6-luna`, `grok-4.6`, `deepseek-v4-pro`, `deepseek-flash`), and the personal-plan routes (`gpt-5.6-terra-codex-subscription`, `gpt-5.6-sol-codex-subscription`, `gpt-5.6-luna-codex-subscription`). For the configured owner, `auto` uses GPT-5.6 Terra through the personal Codex plan; base GPT-5.6 IDs select Azure API, while suffixed IDs explicitly select the personal plan. Never put an image or video model ID in `--agent-model`.
|
|
94
94
|
|
|
95
95
|
```bash
|
|
96
96
|
npx makaron-cli chat --project auto --agent-model deepseek-v4-pro --json -b "make a 20s badminton video"
|
|
@@ -300,11 +300,11 @@ npx makaron-cli edit --image photo.jpg --ref style.jpg "match this style"
|
|
|
300
300
|
# Output to file
|
|
301
301
|
npx makaron-cli edit --image photo.jpg --out result.jpg "make it dramatic"
|
|
302
302
|
|
|
303
|
-
# Strict transparent output through GPT Image 2
|
|
304
|
-
npx makaron-cli edit --image-model
|
|
303
|
+
# Strict transparent output through GPT Image 2.5 Flare
|
|
304
|
+
npx makaron-cli edit --image-model gpt-image-2.5-flare --background transparent --out sticker.png "a magenta star sticker"
|
|
305
305
|
```
|
|
306
306
|
|
|
307
|
-
Options: `--image`, `--image-model gemini|gemini-lite|qwen|openai|wan2.7-image|pony|wai`, `--ref <file>` (up to 3), `--aspect <ratio>`, `--background auto|opaque|transparent`, `--out <path>`. Transparent output routes strictly to GPT Image 2 and fails instead of returning an opaque fallback. Wan 2.7 Image is an explicit fast ~1K route; do not automatically retry failures/timeouts, and do not promise exact face preservation.
|
|
307
|
+
Options: `--image`, `--image-model gemini|gemini-lite|qwen|openai|gpt-image-2.5-flare|gpt-image-2.5-sunburst|wan2.7-image|pony|wai`, `--ref <file>` (up to 3), `--aspect <ratio>`, `--background auto|opaque|transparent`, `--out <path>`. Transparent output routes strictly to GPT Image 2.5 Flare and fails instead of returning an opaque fallback. Wan 2.7 Image is an explicit fast ~1K route; do not automatically retry failures/timeouts, and do not promise exact face preservation.
|
|
308
308
|
|
|
309
309
|
### `video` — Standalone video tools (no project timeline)
|
|
310
310
|
|
|
@@ -490,3 +490,5 @@ send_message "All done!"
|
|
|
490
490
|
- `edit`/`video`/`music` are fallback tools for when `chat` is unavailable or you need raw model access without project context.
|
|
491
491
|
|
|
492
492
|
FAL video models: **fal H3 Turbo** uses `minimax-h3-max` for single-start-frame I2V/T2V. **FAL H3 Max** uses the new selector `fal-h3-max`: native T2V or image/video/audio reference-to-video, default 768p, optional 480p/1080p, integer 5–15s; at most 9 images / 3 videos / 3 audios / 12 total. Reference video/audio each 2–15s and each modality totals at most 15s. Source-video modifications use generation with feature references, not typed edit/extend. Reference input tokens are billed in addition to output video; query current pricing.
|
|
493
|
+
|
|
494
|
+
The legacy `openai` image-model parameter now resolves to GPT Image 2.5 Flare.
|