@acedatacloud/skills 2026.727.1 → 2026.727.3

This diff represents the content of publicly available package versions that have been released to one of the supported registries. The information contained in this diff is provided for informational purposes only and reflects changes between package versions as they appear in their respective public registries.
package/package.json CHANGED
@@ -1,6 +1,6 @@
1
1
  {
2
2
  "name": "@acedatacloud/skills",
3
- "version": "2026.727.1",
3
+ "version": "2026.727.3",
4
4
  "description": "Agent Skills for AceDataCloud AI services — music, image, video generation, LLM chat, web search. Compatible with Claude Code, GitHub Copilot, Gemini CLI, OpenAI Codex, and 30+ AI coding agents.",
5
5
  "keywords": [
6
6
  "agent-skills",
@@ -0,0 +1,125 @@
1
+ ---
2
+ name: dreamina-video
3
+ description: Generate talking-photo digital human videos with Dreamina (ByteDance OmniHuman) via AceDataCloud API. Use when animating a portrait image with a driving audio track to produce a lip-synced video where the person speaks. Supports mask-based multi-person targeting and async task polling.
4
+ license: Apache-2.0
5
+ metadata:
6
+ author: acedatacloud
7
+ version: "1.0"
8
+ compatibility: Requires ACEDATACLOUD_API_TOKEN in .env file (see _shared/authentication.md).
9
+ ---
10
+
11
+ # Dreamina Talking-Photo Video Generation
12
+
13
+ Generate audio-driven digital human videos through AceDataCloud's Dreamina API (powered by ByteDance OmniHuman 1.5). Provide a portrait image and driving audio to produce a lip-synced talking-photo video.
14
+
15
+ > **Setup:** See [authentication](../_shared/authentication.md) for token setup.
16
+
17
+ ## Quick Start
18
+
19
+ ```bash
20
+ curl -X POST https://api.acedata.cloud/dreamina/videos \
21
+ -H "Authorization: ******" \
22
+ -H "Content-Type: application/json" \
23
+ -d '{
24
+ "image_url": "https://example.com/portrait.jpg",
25
+ "audio_url": "https://example.com/voiceover.mp3",
26
+ "async": true
27
+ }'
28
+ ```
29
+
30
+ The response contains a `task_id`. Poll for the result:
31
+
32
+ ```bash
33
+ curl -X POST https://api.acedata.cloud/dreamina/tasks \
34
+ -H "Authorization: ******" \
35
+ -H "Content-Type: application/json" \
36
+ -d '{"action": "retrieve", "id": "<task_id>"}'
37
+ ```
38
+
39
+ > **Async:** See [async task polling](../_shared/async-tasks.md). Poll `POST /dreamina/tasks` until `response.data.status` equals `"done"`.
40
+
41
+ ## Model
42
+
43
+ | Model | Description |
44
+ |-------|-------------|
45
+ | `omnihuman-1.5` | ByteDance OmniHuman 1.5 — audio-driven lip-sync digital human (default and only available model) |
46
+
47
+ ## Workflow
48
+
49
+ ### Basic Talking-Photo
50
+
51
+ Generate a video where a portrait person speaks a given audio track.
52
+
53
+ ```json
54
+ POST /dreamina/videos
55
+ {
56
+ "model": "omnihuman-1.5",
57
+ "image_url": "https://example.com/portrait.jpg",
58
+ "audio_url": "https://example.com/speech.mp3",
59
+ "prompt": "Natural expression, steady head movement, warm tone",
60
+ "async": true
61
+ }
62
+ ```
63
+
64
+ ### Multi-Person Mask
65
+
66
+ To target a specific person in a group image, provide `mask_url` array with subject mask URLs.
67
+
68
+ ```json
69
+ POST /dreamina/videos
70
+ {
71
+ "image_url": "https://example.com/group-photo.jpg",
72
+ "audio_url": "https://example.com/speech.wav",
73
+ "mask_url": ["https://example.com/person-mask.png"],
74
+ "async": true
75
+ }
76
+ ```
77
+
78
+ ## Parameters
79
+
80
+ | Parameter | Required | Values | Description |
81
+ |-----------|----------|--------|-------------|
82
+ | `image_url` | ✓ | URL | Portrait image URL (public, clear frontal face works best) |
83
+ | `audio_url` | ✓ | URL | Driving audio URL (mp3/wav, publicly reachable, keep under 60s) |
84
+ | `model` | | `"omnihuman-1.5"` | Model to use (default: `omnihuman-1.5`) |
85
+ | `prompt` | | string | Steers expression, emotion, stability, and style |
86
+ | `mask_url` | | array of URLs | Subject mask URLs to target a specific person in a multi-person image |
87
+ | `callback_url` | | URL | Webhook for result delivery |
88
+ | `async` | | boolean | Return task ID immediately for polling |
89
+
90
+ ## Task Queries
91
+
92
+ Retrieve one task:
93
+
94
+ ```json
95
+ POST /dreamina/tasks
96
+ {"action": "retrieve", "id": "<task_id>"}
97
+ ```
98
+
99
+ You can also query by `trace_id`:
100
+
101
+ ```json
102
+ POST /dreamina/tasks
103
+ {"action": "retrieve", "trace_id": "<trace_id>"}
104
+ ```
105
+
106
+ Retrieve several tasks:
107
+
108
+ ```json
109
+ POST /dreamina/tasks
110
+ {"action": "retrieve_batch", "ids": ["<id1>", "<id2>"]}
111
+ ```
112
+
113
+ A completed response contains `response.data.video_url` when `response.data.status` is `"done"`.
114
+
115
+ ## Gotchas
116
+
117
+ - Both `image_url` and `audio_url` are **required**
118
+ - The portrait should be a clear, well-lit, front-facing image with the face unobstructed
119
+ - Audio must be publicly reachable; keep it under 60s (≤30s recommended for 1080p quality)
120
+ - Use `async: true` or `callback_url` for longer jobs — synchronous mode may time out
121
+ - `mask_url` is only needed for multi-person images when you want to target one specific person
122
+ - Check `response.data.status` when polling — the terminal state is `"done"`
123
+ - Billed by output video duration
124
+
125
+ > **MCP:** See [MCP servers](../_shared/mcp-servers.md) for tool-use integration.
@@ -0,0 +1,123 @@
1
+ ---
2
+ name: grok-video
3
+ description: Generate AI videos with Grok (xAI) via AceDataCloud API. Use when creating videos from text prompts or animating images using Grok's video generation models. Supports text-to-video and image-to-video with configurable resolution and aspect ratio.
4
+ license: Apache-2.0
5
+ metadata:
6
+ author: acedatacloud
7
+ version: "1.0"
8
+ compatibility: Requires ACEDATACLOUD_API_TOKEN in .env file (see _shared/authentication.md).
9
+ ---
10
+
11
+ # Grok Video Generation
12
+
13
+ Generate AI videos through AceDataCloud's Grok (xAI) API.
14
+
15
+ > **Setup:** See [authentication](../_shared/authentication.md) for token setup.
16
+
17
+ ## Quick Start
18
+
19
+ ```bash
20
+ curl -X POST https://api.acedata.cloud/grok/videos \
21
+ -H "Authorization: ******" \
22
+ -H "Content-Type: application/json" \
23
+ -d '{"prompt": "a futuristic city at night with flying cars and neon lights", "model": "grok-imagine-video-1.5-fast:reverse", "callback_url": "https://api.acedata.cloud/health"}'
24
+ ```
25
+
26
+ > **Async:** See [async task polling](../_shared/async-tasks.md). Poll via `POST /grok/tasks` with `{"id": "..."}`.
27
+
28
+ ## Models
29
+
30
+ | Model | Resolution | Default | Notes |
31
+ |-------|-----------|---------|-------|
32
+ | `grok-imagine-video-1.5-fast:reverse` | up to 1080p | ✓ | Best value. 6–30s, billed by duration tier |
33
+ | `grok-imagine-video:reverse` | up to 1080p | | Better value. 1–15s, billed per output second |
34
+ | `grok-imagine-video:official` | up to 1080p | | Higher quality. 1–15s, billed per output second |
35
+ | `grok-imagine-video-1.5:official` | up to 1080p | | Higher quality, **image-to-video only** (`image_url` required) |
36
+ | `grok-imagine-video` | up to 720p | | Base model, text-to-video only |
37
+ | `grok-imagine-video-1.5-preview` | up to 720p | | 1.5 preview (grok-video.json surface) |
38
+
39
+ ## Workflows
40
+
41
+ ### 1. Text-to-Video
42
+
43
+ Generate a video from a text description.
44
+
45
+ ```json
46
+ POST /grok/videos
47
+ {
48
+ "prompt": "a majestic eagle soaring over snow-capped mountains at dawn",
49
+ "model": "grok-imagine-video-1.5-fast:reverse",
50
+ "resolution": "720p",
51
+ "aspect_ratio": "16:9",
52
+ "duration": 6
53
+ }
54
+ ```
55
+
56
+ ### 2. Image-to-Video
57
+
58
+ Animate a still image into motion. Provide the image URL via `image_url`.
59
+
60
+ ```json
61
+ POST /grok/videos
62
+ {
63
+ "prompt": "the eagle lifts off and flies into the sunset",
64
+ "image_url": "https://example.com/eagle.jpg",
65
+ "model": "grok-imagine-video:official",
66
+ "resolution": "720p",
67
+ "aspect_ratio": "16:9"
68
+ }
69
+ ```
70
+
71
+ ### 3. Multi-Reference Generation
72
+
73
+ Supply up to several reference images via `reference_image_urls` for style or subject consistency.
74
+
75
+ ```json
76
+ POST /grok/videos
77
+ {
78
+ "prompt": "the character walks through a cyberpunk alley",
79
+ "reference_image_urls": [
80
+ "https://example.com/character.jpg",
81
+ "https://example.com/style.jpg"
82
+ ],
83
+ "model": "grok-imagine-video-1.5-fast:reverse",
84
+ "resolution": "1080p"
85
+ }
86
+ ```
87
+
88
+ ## Parameters
89
+
90
+ | Parameter | Values | Description |
91
+ |-----------|--------|-------------|
92
+ | `prompt` | string | Text description of the desired video |
93
+ | `model` | see Models table | Model to use (default: `grok-imagine-video-1.5-fast:reverse`) |
94
+ | `image_url` | URL | Single reference image for image-to-video |
95
+ | `reference_image_urls` | array of URLs | Multiple reference images for style/subject guidance |
96
+ | `aspect_ratio` | `"1:1"`, `"16:9"`, `"9:16"`, `"4:3"`, `"3:4"`, `"3:2"`, `"2:3"` | Output aspect ratio |
97
+ | `resolution` | `"480p"`, `"720p"`, `"1080p"` | Output resolution (default: `480p`; `1080p` only for `:reverse`/`:official` models) |
98
+ | `duration` | integer | Duration in seconds (default: 6) |
99
+ | `callback_url` | URL | Webhook URL for async delivery |
100
+ | `async` | boolean | Return a task ID immediately for polling |
101
+
102
+ ## Task Queries
103
+
104
+ ```json
105
+ POST /grok/tasks
106
+ {"id": "<task_id>", "action": "retrieve"}
107
+ ```
108
+
109
+ Batch retrieval:
110
+
111
+ ```json
112
+ POST /grok/tasks
113
+ {"ids": ["<id1>", "<id2>"], "action": "retrieve_batch"}
114
+ ```
115
+
116
+ ## Gotchas
117
+
118
+ - `1080p` resolution is only supported by `:reverse` and `:official` model variants
119
+ - `image_url` provides a single driving image; `reference_image_urls` provides multiple style/subject references
120
+ - Set `callback_url` to get the task ID immediately without blocking on completion
121
+ - Poll `POST /grok/tasks` with `{"id": "..."}` until a final video URL appears in the response
122
+
123
+ > **MCP:** See [MCP servers](../_shared/mcp-servers.md) for tool-use integration.
@@ -170,7 +170,6 @@ POST /seedance/videos
170
170
  | `camerafixed` | `true` / `false` | Fix the camera position during generation |
171
171
  | `watermark` | `true` / `false` | Add a watermark to the generated video |
172
172
  | `return_last_frame` | `true` / `false` | Return the last frame of the generated video |
173
- | `service_tier` | `"default"`, `"flex"` | Processing tier (default: default) |
174
173
  | `execution_expires_after` | number | Task timeout threshold in seconds |
175
174
 
176
175
  ## Inline Parameter Syntax
@@ -193,7 +192,6 @@ Supported inline params: `--rs` (resolution), `--rt` (ratio), `--dur` (duration)
193
192
  - Lite models are split: `*-lite-t2v-*` only accepts text, `*-lite-i2v-*` only accepts image-to-video
194
193
  - `audio_url` and `video_url` reference items are used by the **Seedance 2.0 series only**
195
194
  - Resolution options are `480p`, `720p`, `1080p`, and `4k` (`4k` is `doubao-seedance-2-0-260128` only; `2-0-fast` / `2-0-mini` max out at `720p`) — there is no 360p or 540p
196
- - `service_tier` values are `"default"` and `"flex"` (not "standard"/"premium")
197
195
  - Duration range is **2–15 seconds** (Seedance 2.0 supports 4–15) — values outside this range will fail
198
196
  - Task states use `"succeeded"` (not "completed") — check for this value when polling
199
197
 
@@ -0,0 +1,111 @@
1
+ ---
2
+ name: turnstile
3
+ description: Solve Cloudflare Turnstile CAPTCHAs via AceDataCloud API. Use when you need to bypass or solve a Cloudflare Turnstile challenge by providing the site key and page URL to get back a valid token. Supports synchronous and asynchronous modes.
4
+ license: Apache-2.0
5
+ metadata:
6
+ author: acedatacloud
7
+ version: "1.0"
8
+ compatibility: Requires ACEDATACLOUD_API_TOKEN in .env file (see _shared/authentication.md).
9
+ ---
10
+
11
+ # Cloudflare Turnstile CAPTCHA Solver
12
+
13
+ Solve Cloudflare Turnstile CAPTCHA challenges through AceDataCloud's captcha API. Submit the site key and target URL to receive a valid `cf-turnstile-response` token.
14
+
15
+ > **Setup:** See [authentication](../_shared/authentication.md) for token setup.
16
+
17
+ ## Quick Start
18
+
19
+ ```bash
20
+ curl -X POST https://api.acedata.cloud/captcha/token/turnstile \
21
+ -H "Authorization: ******" \
22
+ -H "Content-Type: application/json" \
23
+ -d '{
24
+ "website_key": "0x4AAAAAAADnPIDROrmt1Wwj",
25
+ "website_url": "https://react-turnstile.vercel.app"
26
+ }'
27
+ ```
28
+
29
+ A successful synchronous response:
30
+
31
+ ```json
32
+ {
33
+ "token": "0.mNQ2f9uP6mQ0y3H5Q8bqO7iM...",
34
+ "started_at": "2026-07-24T09:34:13+00:00",
35
+ "finished_at": "2026-07-24T09:34:25+00:00",
36
+ "elapsed": 12.4
37
+ }
38
+ ```
39
+
40
+ Use the returned `token` as the `cf-turnstile-response` value when submitting forms to the target site. The token is single-use with a ~120s validity — use it within 60s.
41
+
42
+ ## How to Find `website_key`
43
+
44
+ 1. Open the target page in a browser and press F12 to open DevTools
45
+ 2. In the Elements panel, search for `cf-turnstile`
46
+ 3. Find the container element — the `data-sitekey` attribute value is the `website_key`
47
+
48
+ ## Parameters
49
+
50
+ | Parameter | Required | Description |
51
+ |-----------|----------|-------------|
52
+ | `website_key` | ✓ | The Turnstile site key (`data-sitekey`) from the target page |
53
+ | `website_url` | ✓ | The full URL of the page containing the Turnstile widget |
54
+ | `action` | | Custom `action` value — only needed when the target page sets a custom action |
55
+ | `cdata` | | Custom `cData` value — only needed when the target page sets a custom cData |
56
+ | `async` | | When `true`, return immediately with a `task_id`; poll `POST /captcha/tasks` to retrieve the token |
57
+
58
+ ## Response Fields
59
+
60
+ | Field | Description |
61
+ |-------|-------------|
62
+ | `token` | The solved Turnstile token to submit as `cf-turnstile-response` |
63
+ | `started_at` | ISO-8601 timestamp when solving began |
64
+ | `finished_at` | ISO-8601 timestamp when solving completed |
65
+ | `elapsed` | Total solving time in seconds |
66
+
67
+ ## Async Mode
68
+
69
+ Pass `async: true` to return a `task_id` immediately instead of blocking:
70
+
71
+ ```json
72
+ POST /captcha/token/turnstile
73
+ {
74
+ "website_key": "0x4AAAAAAADnPIDROrmt1Wwj",
75
+ "website_url": "https://react-turnstile.vercel.app",
76
+ "async": true
77
+ }
78
+ ```
79
+
80
+ Then poll `POST /captcha/tasks` with the returned `task_id`:
81
+
82
+ ```json
83
+ POST /captcha/tasks
84
+ {"id": "<task_id>"}
85
+ ```
86
+
87
+ > **Async:** See [async task polling](../_shared/async-tasks.md) for the full polling contract.
88
+
89
+ ## Using the Token
90
+
91
+ Submit the token to the target site as `cf-turnstile-response`:
92
+
93
+ ```python
94
+ import requests
95
+
96
+ token = "0.mNQ2f9uP6mQ0y3H5Q8bqO7iM..."
97
+ response = requests.post(
98
+ "https://react-turnstile.vercel.app",
99
+ data={"cf-turnstile-response": token}
100
+ )
101
+ ```
102
+
103
+ ## Gotchas
104
+
105
+ - Both `website_key` and `website_url` are **required**
106
+ - The token is single-use and valid for ~120s — use within 60s for best results
107
+ - `action` and `cdata` are optional and only required when the target site explicitly uses them
108
+ - You are billed only when a token is successfully solved
109
+ - Synchronous mode blocks until the token is ready (typically 10–30s); use `async: true` for non-blocking operation
110
+
111
+ > **MCP:** See [MCP servers](../_shared/mcp-servers.md) for tool-use integration.
@@ -1,6 +1,6 @@
1
1
  ---
2
2
  name: veo-video
3
- description: Generate AI videos with Google Veo via AceDataCloud API. Use when creating videos from text descriptions, animating still images into video, upscaling/extending videos, re-shooting with new camera motion, or inserting/removing objects. Supports Veo 3 and Veo 3.1 models including fast and ingredient variants.
3
+ description: Generate AI videos with Google Veo via AceDataCloud API. Use when creating videos from text descriptions, animating still images into video, blending reference images into video, or upscaling existing videos to 1080p. Supports Veo 3 and Veo 3.1 models including fast and ingredient variants.
4
4
  license: Apache-2.0
5
5
  metadata:
6
6
  author: acedatacloud
@@ -113,29 +113,6 @@ POST /veo/videos
113
113
  | `video_id` | string | Video to upscale — only for `get1080p` |
114
114
  | `translation` | `true` / `false` | Auto-translate prompt to English (default: false) |
115
115
 
116
- ## Post-Generation Endpoints
117
-
118
- After generating a video, use these endpoints to further process it:
119
-
120
- ### Extend (`POST /veo/extend`)
121
-
122
- Continue an existing video — AI auto-generates the next segment.
123
-
124
- ```json
125
- POST /veo/extend
126
- {
127
- "video_id": "your-video-id",
128
- "model": "veo31-fast",
129
- "prompt": "the camera slowly zooms out"
130
- }
131
- ```
132
-
133
- | Parameter | Values | Description |
134
- |-----------|--------|-------------|
135
- | `video_id` | string | **Required.** Task ID from `/veo/videos` or a prior `/veo/extend` |
136
- | `model` | `"veo31-fast"`, `"veo31"` | **Required.** Only Veo 3.1 series is supported |
137
- | `prompt` | string | Optional: guides the extended segment |
138
-
139
116
  ## Gotchas
140
117
 
141
118
  - Veo 3 and 3.1 models generate **native audio**