makaron-cli 0.14.8 → 0.14.10

This diff represents the content of publicly available package versions that have been released to one of the supported registries. The information contained in this diff is provided for informational purposes only and reflects changes between package versions as they appear in their respective public registries.
@@ -1,6 +1,6 @@
1
1
  {
2
2
  "name": "makaron-cli",
3
- "version": "0.14.8",
3
+ "version": "0.14.9",
4
4
  "description": "Give Claude Code a creative agent. Pass complete creative requests and source media to Makaron Chat.",
5
5
  "author": {
6
6
  "name": "Versa AI",
@@ -1,6 +1,6 @@
1
1
  {
2
2
  "name": "makaron-cli",
3
- "version": "0.14.8",
3
+ "version": "0.14.9",
4
4
  "description": "Give Codex a creative agent. Pass complete creative requests and source media to Makaron Chat.",
5
5
  "author": {
6
6
  "name": "Versa AI",
package/README.md CHANGED
@@ -98,7 +98,7 @@ npx makaron-cli chat --project auto --image photo.jpg --json -b "make it cinemat
98
98
  npx makaron-cli chat --project auto --image img1.jpg --image img2.jpg --json -b "combine these"
99
99
  ```
100
100
 
101
- `chat` routes image and video models automatically, but you may select the Agent LLM with `--agent-model`. Accepted values are `auto`, the base model IDs (`gpt-5.6-terra`, `gpt-5.6-sol`, `gpt-5.6-luna`, `grok-4.6`, `deepseek-v4-pro`), and the personal-plan routes (`gpt-5.6-terra-codex-subscription`, `gpt-5.6-sol-codex-subscription`, `gpt-5.6-luna-codex-subscription`, `grok-4.6-grok-subscription`). For the configured owner, `auto` resolves to GPT-5.6 Terra through the personal Codex plan. Base GPT-5.6 IDs select Azure API and base `grok-4.6` selects OpenRouter API; the suffixed IDs select the corresponding personal plan explicitly. This flag changes only the reasoning/tool-calling Agent LLM.
101
+ `chat` routes image and video models automatically, but you may select the Agent LLM with `--agent-model`. Accepted values are `auto`, the base model IDs (`gpt-5.6-terra`, `gpt-5.6-sol`, `gpt-5.6-luna`, `grok-4.6`, `deepseek-v4-pro`, `deepseek-flash`), and the personal-plan routes (`gpt-5.6-terra-codex-subscription`, `gpt-5.6-sol-codex-subscription`, `gpt-5.6-luna-codex-subscription`, `grok-4.6-grok-subscription`). For the configured owner, `auto` resolves to GPT-5.6 Terra through the personal Codex plan. Base GPT-5.6 IDs select Azure API and base `grok-4.6` selects OpenRouter API; the suffixed IDs select the corresponding personal plan explicitly. This flag changes only the reasoning/tool-calling Agent LLM.
102
102
 
103
103
  ```bash
104
104
  # Explicit lower-cost Agent LLM for a controlled comparison
@@ -363,11 +363,11 @@ npx makaron-cli edit --image photo.jpg --ref style.jpg "match this style"
363
363
  # Output to file
364
364
  npx makaron-cli edit --image photo.jpg --out result.jpg "make it dramatic"
365
365
 
366
- # Strict transparent PNG/WebP output through GPT Image 2 (fails rather than returning opaque)
367
- npx makaron-cli edit --image-model openai --background transparent --out sticker.png "a magenta star sticker"
366
+ # Strict transparent PNG/WebP output through GPT Image 2.5 Flare (fails rather than returning opaque)
367
+ npx makaron-cli edit --image-model gpt-image-2.5-flare --background transparent --out sticker.png "a magenta star sticker"
368
368
  ```
369
369
 
370
- Options: `--image`, `--image-model gemini|gemini-lite|qwen|openai|wan2.7-image|pony|wai`, `--ref <file>` (up to 3), `--aspect <ratio>`, `--background auto|opaque|transparent`, `--out <path>`. Transparent output routes strictly to GPT Image 2 and is returned only when the provider supplies real PNG/WebP alpha.
370
+ Options: `--image`, `--image-model gemini|gemini-lite|qwen|openai|gpt-image-2.5-flare|gpt-image-2.5-sunburst|wan2.7-image|pony|wai`, `--ref <file>` (up to 3), `--aspect <ratio>`, `--background auto|opaque|transparent`, `--out <path>`. Transparent output routes strictly to GPT Image 2.5 Flare and is returned only when the provider supplies real PNG/WebP alpha.
371
371
 
372
372
  `wan2.7-image` uses Alibaba international for fast, approximately 1K generation and editing (default 6 credits/image). Failed or timed-out Wan requests are not automatically retried or switched to another model. Face identity can change. Example: `makaron edit --image portrait.jpg --image-model wan2.7-image --aspect 16:9 --out stadium.jpg "Place this woman in a baseball stadium, preserving her face."`
373
373
 
@@ -656,3 +656,5 @@ npx makaron-cli admin fetch-skill https://www.makaron.app/s/4c4cbd57
656
656
  - Before photo should match the person in the cover (hair, clothing, accessories)
657
657
 
658
658
  FAL video models: **fal H3 Turbo** uses `minimax-h3-max` for single-start-frame I2V/T2V. **FAL H3 Max** uses the new selector `fal-h3-max`: native T2V or image/video/audio reference-to-video, default 768p, optional 480p/1080p, integer 5–15s; at most 9 images / 3 videos / 3 audios / 12 total. Reference video/audio each 2–15s and each modality totals at most 15s. Source-video modifications use generation with feature references, not typed edit/extend. Reference input tokens are billed in addition to output video; query current pricing.
659
+
660
+ The legacy `openai` image-model parameter now resolves to GPT Image 2.5 Flare.
package/bin/makaron.mjs CHANGED
@@ -39,6 +39,7 @@ const CHAT_AGENT_MODELS = [
39
39
  'grok-4.6',
40
40
  'grok-4.6-grok-subscription',
41
41
  'deepseek-v4-pro',
42
+ 'deepseek-flash',
42
43
  ];
43
44
 
44
45
  // Public anon key (safe to embed — only enables auth, not data access)
@@ -462,7 +463,7 @@ Options:
462
463
  --media-manifest <file|-> Import typed image/video media before this run.
463
464
  --skill <id|label|name> Use an installed skill or auto-install a matched marketplace skill.
464
465
  --agent-model <id> Agent LLM only: auto, gpt-5.6-terra, gpt-5.6-sol,
465
- gpt-5.6-luna, grok-4.6, deepseek-v4-pro, or a
466
+ gpt-5.6-luna, grok-4.6, deepseek-v4-pro, deepseek-flash, or a
466
467
  gpt-5.6-*-codex-subscription or
467
468
  grok-4.6-grok-subscription personal-plan route.
468
469
  --background, -b Submit and print a runId.
@@ -1879,7 +1880,7 @@ Usage:
1879
1880
  Options:
1880
1881
  --image <file|url> Base image to edit. Omit for text-to-image.
1881
1882
  --ref <file|url> Additional reference image. Repeatable, up to 3.
1882
- --image-model <id> gemini, gemini-lite, qwen, openai, wan2.7-image, pony, or wai.
1883
+ --image-model <id> gemini, gemini-lite, qwen, openai, gpt-image-2.5-flare, gpt-image-2.5-sunburst, wan2.7-image, pony, or wai.
1883
1884
  --skill <id> enhance, creative, wild, or captions.
1884
1885
  --aspect <ratio> Output aspect ratio, for example 1:1, 16:9, or 9:16.
1885
1886
  --background <mode> auto, opaque, or transparent.
@@ -1887,13 +1888,13 @@ Options:
1887
1888
  --help, -h Show this help.
1888
1889
 
1889
1890
  Notes:
1890
- Model selection is optional. Transparent output routes strictly to GPT Image 2
1891
+ Model selection is optional. Transparent output routes strictly to GPT Image 2.5 Flare
1891
1892
  and fails instead of returning an opaque fallback.
1892
1893
 
1893
1894
  Examples:
1894
1895
  makaron edit --image portrait.jpg --image-model qwen --out result.jpg "cinematic warm light"
1895
1896
  makaron edit --image product.jpg --ref style.png --aspect 1:1 "use this visual style"
1896
- makaron edit --image-model openai --background transparent --out sticker.png "a magenta star sticker"
1897
+ makaron edit --image-model gpt-image-2.5-flare --background transparent --out sticker.png "a magenta star sticker"
1897
1898
  `);
1898
1899
  }
1899
1900
 
@@ -2930,7 +2931,7 @@ if (!command || command === '--help' || command === '-h' || command === 'help')
2930
2931
  else promptParts.push(args[i]);
2931
2932
  }
2932
2933
  editArgs.editPrompt = promptParts.join(' ');
2933
- if (!editArgs.editPrompt) { console.error('Usage: makaron edit [--image <file|url>] [--image-model gemini|gemini-lite|qwen|openai|wan2.7-image|pony|wai] [--ref <file>] [--aspect <ratio>] [--background auto|opaque|transparent] [--out <file>] "prompt"'); process.exit(1); }
2934
+ if (!editArgs.editPrompt) { console.error('Usage: makaron edit [--image <file|url>] [--image-model gemini|gemini-lite|qwen|openai|gpt-image-2.5-flare|gpt-image-2.5-sunburst|wan2.7-image|pony|wai] [--ref <file>] [--aspect <ratio>] [--background auto|opaque|transparent] [--out <file>] "prompt"'); process.exit(1); }
2934
2935
  process.stderr.write('🎨 Generating...\n');
2935
2936
  const result = await callMcpTool(baseUrl, headers, 'makaron_edit_image', editArgs);
2936
2937
  saveMcpImage(result, outputPath);
package/package.json CHANGED
@@ -1,6 +1,6 @@
1
1
  {
2
2
  "name": "makaron-cli",
3
- "version": "0.14.8",
3
+ "version": "0.14.10",
4
4
  "description": "Talk to Makaron Agent from the terminal — create projects, edit images, generate videos",
5
5
  "type": "module",
6
6
  "scripts": {
@@ -90,7 +90,7 @@ npx makaron-cli chat --project auto --image photo.jpg --json -b "make it cinemat
90
90
  npx makaron-cli chat --project auto --image img1.jpg --image img2.jpg --json -b "combine these"
91
91
  ```
92
92
 
93
- `chat` routes image and video models automatically. Use `--agent-model` only when the user explicitly asks to select or compare the reasoning/tool-calling Agent LLM. Accepted values are `auto`, the base model IDs (`gpt-5.6-terra`, `gpt-5.6-sol`, `gpt-5.6-luna`, `grok-4.6`, `deepseek-v4-pro`), and the personal-plan routes (`gpt-5.6-terra-codex-subscription`, `gpt-5.6-sol-codex-subscription`, `gpt-5.6-luna-codex-subscription`). For the configured owner, `auto` uses GPT-5.6 Terra through the personal Codex plan; base GPT-5.6 IDs select Azure API, while suffixed IDs explicitly select the personal plan. Never put an image or video model ID in `--agent-model`.
93
+ `chat` routes image and video models automatically. Use `--agent-model` only when the user explicitly asks to select or compare the reasoning/tool-calling Agent LLM. Accepted values are `auto`, the base model IDs (`gpt-5.6-terra`, `gpt-5.6-sol`, `gpt-5.6-luna`, `grok-4.6`, `deepseek-v4-pro`, `deepseek-flash`), and the personal-plan routes (`gpt-5.6-terra-codex-subscription`, `gpt-5.6-sol-codex-subscription`, `gpt-5.6-luna-codex-subscription`). For the configured owner, `auto` uses GPT-5.6 Terra through the personal Codex plan; base GPT-5.6 IDs select Azure API, while suffixed IDs explicitly select the personal plan. Never put an image or video model ID in `--agent-model`.
94
94
 
95
95
  ```bash
96
96
  npx makaron-cli chat --project auto --agent-model deepseek-v4-pro --json -b "make a 20s badminton video"
@@ -300,11 +300,11 @@ npx makaron-cli edit --image photo.jpg --ref style.jpg "match this style"
300
300
  # Output to file
301
301
  npx makaron-cli edit --image photo.jpg --out result.jpg "make it dramatic"
302
302
 
303
- # Strict transparent output through GPT Image 2
304
- npx makaron-cli edit --image-model openai --background transparent --out sticker.png "a magenta star sticker"
303
+ # Strict transparent output through GPT Image 2.5 Flare
304
+ npx makaron-cli edit --image-model gpt-image-2.5-flare --background transparent --out sticker.png "a magenta star sticker"
305
305
  ```
306
306
 
307
- Options: `--image`, `--image-model gemini|gemini-lite|qwen|openai|wan2.7-image|pony|wai`, `--ref <file>` (up to 3), `--aspect <ratio>`, `--background auto|opaque|transparent`, `--out <path>`. Transparent output routes strictly to GPT Image 2 and fails instead of returning an opaque fallback. Wan 2.7 Image is an explicit fast ~1K route; do not automatically retry failures/timeouts, and do not promise exact face preservation.
307
+ Options: `--image`, `--image-model gemini|gemini-lite|qwen|openai|gpt-image-2.5-flare|gpt-image-2.5-sunburst|wan2.7-image|pony|wai`, `--ref <file>` (up to 3), `--aspect <ratio>`, `--background auto|opaque|transparent`, `--out <path>`. Transparent output routes strictly to GPT Image 2.5 Flare and fails instead of returning an opaque fallback. Wan 2.7 Image is an explicit fast ~1K route; do not automatically retry failures/timeouts, and do not promise exact face preservation.
308
308
 
309
309
  ### `video` — Standalone video tools (no project timeline)
310
310
 
@@ -490,3 +490,5 @@ send_message "All done!"
490
490
  - `edit`/`video`/`music` are fallback tools for when `chat` is unavailable or you need raw model access without project context.
491
491
 
492
492
  FAL video models: **fal H3 Turbo** uses `minimax-h3-max` for single-start-frame I2V/T2V. **FAL H3 Max** uses the new selector `fal-h3-max`: native T2V or image/video/audio reference-to-video, default 768p, optional 480p/1080p, integer 5–15s; at most 9 images / 3 videos / 3 audios / 12 total. Reference video/audio each 2–15s and each modality totals at most 15s. Source-video modifications use generation with feature references, not typed edit/extend. Reference input tokens are billed in addition to output video; query current pricing.
493
+
494
+ The legacy `openai` image-model parameter now resolves to GPT Image 2.5 Flare.