@mastra/mcp-docs-server 1.2.27-alpha.15 → 1.2.27-alpha.19
This diff represents the content of publicly available package versions that have been released to one of the supported registries. The information contained in this diff is provided for informational purposes only and reflects changes between package versions as they appear in their respective public registries.
- package/.docs/models/environment-variables.md +211 -211
- package/.docs/models/gateways/openrouter.md +2 -1
- package/.docs/models/gateways/vercel.md +4 -1
- package/.docs/models/index.md +1 -1
- package/.docs/models/providers/kilo.md +9 -9
- package/.docs/models/providers/llmgateway-providers.md +1 -3
- package/.docs/models/providers/minimax-cn-coding-plan.md +5 -5
- package/.docs/models/providers/minimax-cn.md +5 -5
- package/.docs/models/providers/nano-gpt.md +9 -4
- package/.docs/models/providers/nebius.md +4 -1
- package/.docs/models/providers/opencode.md +5 -1
- package/.docs/models/providers/tensorx.md +31 -32
- package/.docs/models/providers/zai.md +3 -2
- package/.docs/models/providers/zenmux.md +21 -14
- package/.docs/models/providers/zhipuai.md +3 -2
- package/.docs/models/providers.md +2 -2
- package/.docs/reference/agents/generate.md +1 -1
- package/.docs/reference/agents/listSuspendedRuns.md +2 -0
- package/.docs/reference/agents/network.md +1 -1
- package/.docs/reference/processors/provider-history-compat.md +17 -8
- package/.docs/reference/streaming/agents/stream.md +1 -1
- package/package.json +3 -3
|
@@ -4,7 +4,7 @@
|
|
|
4
4
|
|
|
5
5
|
# OpenRouter
|
|
6
6
|
|
|
7
|
-
OpenRouter aggregates models from multiple providers with enhanced features like rate limiting and failover. Access
|
|
7
|
+
OpenRouter aggregates models from multiple providers with enhanced features like rate limiting and failover. Access 372 models through Mastra's model router.
|
|
8
8
|
|
|
9
9
|
Learn more in the [OpenRouter documentation](https://openrouter.ai/models).
|
|
10
10
|
|
|
@@ -408,4 +408,5 @@ ANTHROPIC_API_KEY=ant-...
|
|
|
408
408
|
| `z-ai/glm-5.2:free` |
|
|
409
409
|
| `z-ai/glm-5.3` |
|
|
410
410
|
| `z-ai/glm-5.3-flash` |
|
|
411
|
+
| `z-ai/glm-5.3-flashx` |
|
|
411
412
|
| `z-ai/glm-5v-turbo` |
|
|
@@ -4,7 +4,7 @@
|
|
|
4
4
|
|
|
5
5
|
# Vercel
|
|
6
6
|
|
|
7
|
-
Vercel aggregates models from multiple providers with enhanced features like rate limiting and failover. Access
|
|
7
|
+
Vercel aggregates models from multiple providers with enhanced features like rate limiting and failover. Access 375 models through Mastra's model router.
|
|
8
8
|
|
|
9
9
|
Learn more in the [Vercel documentation](https://ai-sdk.dev/providers/ai-sdk-providers).
|
|
10
10
|
|
|
@@ -234,6 +234,7 @@ ANTHROPIC_API_KEY=ant-...
|
|
|
234
234
|
| `mistral/mistral-medium-3.5` |
|
|
235
235
|
| `mistral/mistral-nemo` |
|
|
236
236
|
| `mistral/mistral-small` |
|
|
237
|
+
| `mixedbread/toast-1` |
|
|
237
238
|
| `moonshotai/kimi-k2` |
|
|
238
239
|
| `moonshotai/kimi-k2-thinking` |
|
|
239
240
|
| `moonshotai/kimi-k2.5` |
|
|
@@ -337,6 +338,8 @@ ANTHROPIC_API_KEY=ant-...
|
|
|
337
338
|
| `poolside/laguna-s-2.1-free` |
|
|
338
339
|
| `prodia/flux-fast-schnell` |
|
|
339
340
|
| `quiverai/arrow-1.1` |
|
|
341
|
+
| `quiverai/arrow-2` |
|
|
342
|
+
| `quiverai/arrow-2-telos` |
|
|
340
343
|
| `recraft/recraft-v2` |
|
|
341
344
|
| `recraft/recraft-v3` |
|
|
342
345
|
| `recraft/recraft-v4` |
|
package/.docs/models/index.md
CHANGED
|
@@ -4,7 +4,7 @@
|
|
|
4
4
|
|
|
5
5
|
# Model Providers
|
|
6
6
|
|
|
7
|
-
Mastra provides a unified interface for working with LLMs across multiple providers, giving you access to
|
|
7
|
+
Mastra provides a unified interface for working with LLMs across multiple providers, giving you access to 7369 models from 209 providers through a single API.
|
|
8
8
|
|
|
9
9
|
## Features
|
|
10
10
|
|
|
@@ -42,12 +42,12 @@ for await (const chunk of stream) {
|
|
|
42
42
|
| `kilo/~anthropic/claude-haiku-latest` | 200K | | | | | | $1 | $5 |
|
|
43
43
|
| `kilo/~anthropic/claude-opus-latest` | 1.0M | | | | | | $5 | $25 |
|
|
44
44
|
| `kilo/~anthropic/claude-sonnet-latest` | 1.0M | | | | | | $2 | $10 |
|
|
45
|
-
| `kilo/~deepseek/deepseek-flash-latest` | 1.0M | | | | | | $0.
|
|
45
|
+
| `kilo/~deepseek/deepseek-flash-latest` | 1.0M | | | | | | $0.13 | $0.52 |
|
|
46
46
|
| `kilo/~deepseek/deepseek-pro-latest` | 1.0M | | | | | | $0.58 | $2 |
|
|
47
|
-
| `kilo/~deepseek/deepseek-v4-flash-latest` | 1.0M | | | | | | $0.
|
|
47
|
+
| `kilo/~deepseek/deepseek-v4-flash-latest` | 1.0M | | | | | | $0.04 | $0.08 |
|
|
48
48
|
| `kilo/~google/gemini-flash-latest` | 1.0M | | | | | | $0.75 | $4 |
|
|
49
49
|
| `kilo/~google/gemini-pro-latest` | 1.0M | | | | | | $2 | $12 |
|
|
50
|
-
| `kilo/~moonshotai/kimi-latest` | 1.0M | | | | | | $2 | $
|
|
50
|
+
| `kilo/~moonshotai/kimi-latest` | 1.0M | | | | | | $2 | $9 |
|
|
51
51
|
| `kilo/~openai/gpt-astra-latest` | 1.1M | | | | | | $10 | $50 |
|
|
52
52
|
| `kilo/~openai/gpt-luna-latest` | 1.1M | | | | | | $0.20 | $1 |
|
|
53
53
|
| `kilo/~openai/gpt-mini-latest` | 400K | | | | | | $0.75 | $5 |
|
|
@@ -55,7 +55,7 @@ for await (const chunk of stream) {
|
|
|
55
55
|
| `kilo/~openai/gpt-terra-latest` | 1.1M | | | | | | $2 | $12 |
|
|
56
56
|
| `kilo/~x-ai/grok-latest` | 500K | | | | | | $2 | $6 |
|
|
57
57
|
| `kilo/~z-ai/glm-flash-latest` | 1.0M | | | | | | $0.07 | $0.25 |
|
|
58
|
-
| `kilo/~z-ai/glm-latest` | 1.0M | | | | | | $0.
|
|
58
|
+
| `kilo/~z-ai/glm-latest` | 1.0M | | | | | | $0.84 | $3 |
|
|
59
59
|
| `kilo/aion-labs/aion-2.0` | 131K | | | | | | $0.80 | $2 |
|
|
60
60
|
| `kilo/aion-labs/aion-3.0` | 131K | | | | | | $3 | $6 |
|
|
61
61
|
| `kilo/aion-labs/aion-3.0-mini` | 131K | | | | | | $0.70 | $1 |
|
|
@@ -212,7 +212,7 @@ for await (const chunk of stream) {
|
|
|
212
212
|
| `kilo/moonshotai/kimi-k2.5` | 262K | | | | | | $0.60 | $3 |
|
|
213
213
|
| `kilo/moonshotai/kimi-k2.6` | 262K | | | | | | $0.80 | $3 |
|
|
214
214
|
| `kilo/moonshotai/kimi-k2.7-code` | 262K | | | | | | $0.95 | $4 |
|
|
215
|
-
| `kilo/moonshotai/kimi-k3` | 1.0M | | | | | | $2 | $
|
|
215
|
+
| `kilo/moonshotai/kimi-k3` | 1.0M | | | | | | $2 | $9 |
|
|
216
216
|
| `kilo/morph/morph-v3-fast` | 82K | | | | | | $0.80 | $1 |
|
|
217
217
|
| `kilo/morph/morph-v3-large` | 262K | | | | | | $0.90 | $2 |
|
|
218
218
|
| `kilo/nex-agi/nex-n2.5-mini:free` | 262K | | | | | | — | — |
|
|
@@ -224,11 +224,11 @@ for await (const chunk of stream) {
|
|
|
224
224
|
| `kilo/nvidia/nemotron-3-nano-omni-30b-a3b-reasoning:free` | 256K | | | | | | — | — |
|
|
225
225
|
| `kilo/nvidia/nemotron-3-super-120b-a12b` | 262K | | | | | | $0.08 | $0.45 |
|
|
226
226
|
| `kilo/nvidia/nemotron-3-super-120b-a12b:free` | 262K | | | | | | — | — |
|
|
227
|
-
| `kilo/nvidia/nemotron-3-ultra-550b-a55b` |
|
|
227
|
+
| `kilo/nvidia/nemotron-3-ultra-550b-a55b` | 203K | | | | | | $0.50 | $2 |
|
|
228
228
|
| `kilo/nvidia/nemotron-3-ultra-550b-a55b:free` | 1.0M | | | | | | — | — |
|
|
229
229
|
| `kilo/nvidia/nemotron-3.5-content-safety` | 131K | | | | | | $0.20 | $0.20 |
|
|
230
230
|
| `kilo/nvidia/nemotron-3.5-content-safety:free` | 128K | | | | | | — | — |
|
|
231
|
-
| `kilo/nvidia/nemotron-3.5-lightning` | 262K | | | | | | $0.
|
|
231
|
+
| `kilo/nvidia/nemotron-3.5-lightning` | 262K | | | | | | $0.04 | $0.18 |
|
|
232
232
|
| `kilo/nvidia/nemotron-3.5-lightning:free` | 1.0M | | | | | | — | — |
|
|
233
233
|
| `kilo/openai/gpt-3.5-turbo` | 16K | | | | | | $0.50 | $2 |
|
|
234
234
|
| `kilo/openai/gpt-3.5-turbo-0613` | 4K | | | | | | $1 | $2 |
|
|
@@ -270,7 +270,6 @@ for await (const chunk of stream) {
|
|
|
270
270
|
| `kilo/openai/gpt-5.6-luna` | 1.1M | | | | | | $0.20 | $1 |
|
|
271
271
|
| `kilo/openai/gpt-5.6-luna-pro` | 1.1M | | | | | | $0.20 | $1 |
|
|
272
272
|
| `kilo/openai/gpt-5.6-sol` | 1.1M | | | | | | $4 | $20 |
|
|
273
|
-
| `kilo/openai/gpt-5.6-sol-discounted` | 1.1M | | | | | | $2 | $10 |
|
|
274
273
|
| `kilo/openai/gpt-5.6-sol-pro` | 1.1M | | | | | | $4 | $20 |
|
|
275
274
|
| `kilo/openai/gpt-5.6-terra` | 1.1M | | | | | | $2 | $12 |
|
|
276
275
|
| `kilo/openai/gpt-5.6-terra-pro` | 1.1M | | | | | | $2 | $12 |
|
|
@@ -280,7 +279,7 @@ for await (const chunk of stream) {
|
|
|
280
279
|
| `kilo/openai/gpt-audio-mini` | 128K | | | | | | $0.60 | $2 |
|
|
281
280
|
| `kilo/openai/gpt-chat-latest` | 400K | | | | | | $5 | $30 |
|
|
282
281
|
| `kilo/openai/gpt-oss-120b` | 131K | | | | | | $0.03 | $0.17 |
|
|
283
|
-
| `kilo/openai/gpt-oss-20b` | 131K | | | | | | $0.02 | $0.
|
|
282
|
+
| `kilo/openai/gpt-oss-20b` | 131K | | | | | | $0.02 | $0.09 |
|
|
284
283
|
| `kilo/openai/gpt-oss-safeguard-20b` | 131K | | | | | | $0.07 | $0.30 |
|
|
285
284
|
| `kilo/openai/o1` | 200K | | | | | | $15 | $60 |
|
|
286
285
|
| `kilo/openai/o1-pro` | 200K | | | | | | $150 | $600 |
|
|
@@ -416,6 +415,7 @@ for await (const chunk of stream) {
|
|
|
416
415
|
| `kilo/z-ai/glm-5.2:free` | 33K | | | | | | — | — |
|
|
417
416
|
| `kilo/z-ai/glm-5.3` | 1.0M | | | | | | $1 | $4 |
|
|
418
417
|
| `kilo/z-ai/glm-5.3-flash` | 1.0M | | | | | | $0.15 | $0.50 |
|
|
418
|
+
| `kilo/z-ai/glm-5.3-flashx` | 1.0M | | | | | | $0.37 | $1 |
|
|
419
419
|
| `kilo/z-ai/glm-5v-turbo` | 203K | | | | | | $1 | $4 |
|
|
420
420
|
|
|
421
421
|
Model availability, capabilities, context windows, and pricing are sourced from [models.dev](https://models.dev) and may change.
|
|
@@ -4,7 +4,7 @@
|
|
|
4
4
|
|
|
5
5
|
# LLM Gateway
|
|
6
6
|
|
|
7
|
-
Access
|
|
7
|
+
Access 402 LLM Gateway models through Mastra's model router. Authentication is handled automatically using the `LLMGATEWAY_API_KEY` environment variable.
|
|
8
8
|
|
|
9
9
|
Learn more in the [LLM Gateway documentation](https://llmgateway.io/docs).
|
|
10
10
|
|
|
@@ -415,8 +415,6 @@ for await (const chunk of stream) {
|
|
|
415
415
|
| `llmgateway-providers/vertex-openai/qwen3-coder-480b-a35b-instruct` | 262K | | | | | | $0.22 | $2 |
|
|
416
416
|
| `llmgateway-providers/vertex-openai/qwen3-next-80b-a3b-instruct` | 131K | | | | | | $0.15 | $1 |
|
|
417
417
|
| `llmgateway-providers/vertex-openai/qwen3-next-80b-a3b-thinking` | 131K | | | | | | $0.15 | $1 |
|
|
418
|
-
| `llmgateway-providers/vichar-ai/glm-5.3` | 1.0M | | | | | | $1 | $4 |
|
|
419
|
-
| `llmgateway-providers/vichar-ai/glm-5.3-flash` | 1.0M | | | | | | $0.15 | $0.50 |
|
|
420
418
|
| `llmgateway-providers/xai/grok-4` | 256K | | | | | | $3 | $15 |
|
|
421
419
|
| `llmgateway-providers/xai/grok-4-20-beta-0309-non-reasoning` | 2.0M | | | | | | $2 | $6 |
|
|
422
420
|
| `llmgateway-providers/xai/grok-4-20-beta-0309-reasoning` | 2.0M | | | | | | $2 | $6 |
|
|
@@ -2,11 +2,11 @@
|
|
|
2
2
|
|
|
3
3
|
> Discover all available pages from the documentation index: https://mastra.ai/llms.txt
|
|
4
4
|
|
|
5
|
-
# MiniMax Token Plan (minimax.cn)
|
|
6
6
|
|
|
7
|
-
Access 7 MiniMax Token Plan (
|
|
7
|
+
Access 7 MiniMax Token Plan (minimax.cn) models through Mastra's model router. Authentication is handled automatically using the `MINIMAX_API_KEY` environment variable.
|
|
8
8
|
|
|
9
|
-
Learn more in the [MiniMax Token Plan (
|
|
9
|
+
Learn more in the [MiniMax Token Plan (minimax.cn) documentation](https://platform.minimaxi.com/docs/token-plan/intro).
|
|
10
10
|
|
|
11
11
|
```bash
|
|
12
12
|
MINIMAX_API_KEY=your-api-key
|
|
@@ -32,7 +32,7 @@ for await (const chunk of stream) {
|
|
|
32
32
|
}
|
|
33
33
|
```
|
|
34
34
|
|
|
35
|
-
> **Note:** Mastra uses the OpenAI-compatible `/chat/completions` endpoint. Some provider-specific features may not be available. Check the [MiniMax Token Plan (
|
|
35
|
+
> **Note:** Mastra uses the OpenAI-compatible `/chat/completions` endpoint. Some provider-specific features may not be available. Check the [MiniMax Token Plan (minimax.cn) documentation](https://platform.minimaxi.com/docs/token-plan/intro) for details.
|
|
36
36
|
|
|
37
37
|
## Models
|
|
38
38
|
|
|
@@ -57,7 +57,7 @@ const agent = new Agent({
|
|
|
57
57
|
id: "custom-agent",
|
|
58
58
|
name: "custom-agent",
|
|
59
59
|
model: {
|
|
60
|
-
url: "https://api.
|
|
60
|
+
url: "https://api.minimax.cn/anthropic/v1",
|
|
61
61
|
id: "minimax-cn-coding-plan/MiniMax-M2",
|
|
62
62
|
apiKey: process.env.MINIMAX_API_KEY,
|
|
63
63
|
headers: {
|
|
@@ -2,11 +2,11 @@
|
|
|
2
2
|
|
|
3
3
|
> Discover all available pages from the documentation index: https://mastra.ai/llms.txt
|
|
4
4
|
|
|
5
|
-
# MiniMax (minimax.cn)
|
|
6
6
|
|
|
7
|
-
Access 7 MiniMax (
|
|
7
|
+
Access 7 MiniMax (minimax.cn) models through Mastra's model router. Authentication is handled automatically using the `MINIMAX_API_KEY` environment variable.
|
|
8
8
|
|
|
9
|
-
Learn more in the [MiniMax (
|
|
9
|
+
Learn more in the [MiniMax (minimax.cn) documentation](https://platform.minimaxi.com/docs/guides/quickstart).
|
|
10
10
|
|
|
11
11
|
```bash
|
|
12
12
|
MINIMAX_API_KEY=your-api-key
|
|
@@ -32,7 +32,7 @@ for await (const chunk of stream) {
|
|
|
32
32
|
}
|
|
33
33
|
```
|
|
34
34
|
|
|
35
|
-
> **Note:** Mastra uses the OpenAI-compatible `/chat/completions` endpoint. Some provider-specific features may not be available. Check the [MiniMax (
|
|
35
|
+
> **Note:** Mastra uses the OpenAI-compatible `/chat/completions` endpoint. Some provider-specific features may not be available. Check the [MiniMax (minimax.cn) documentation](https://platform.minimaxi.com/docs/guides/quickstart) for details.
|
|
36
36
|
|
|
37
37
|
## Models
|
|
38
38
|
|
|
@@ -57,7 +57,7 @@ const agent = new Agent({
|
|
|
57
57
|
id: "custom-agent",
|
|
58
58
|
name: "custom-agent",
|
|
59
59
|
model: {
|
|
60
|
-
url: "https://api.
|
|
60
|
+
url: "https://api.minimax.cn/anthropic/v1",
|
|
61
61
|
id: "minimax-cn/MiniMax-M2",
|
|
62
62
|
apiKey: process.env.MINIMAX_API_KEY,
|
|
63
63
|
headers: {
|
|
@@ -4,7 +4,7 @@
|
|
|
4
4
|
|
|
5
5
|
# NanoGPT
|
|
6
6
|
|
|
7
|
-
Access
|
|
7
|
+
Access 577 NanoGPT models through Mastra's model router. Authentication is handled automatically using the `NANO_GPT_API_KEY` environment variable.
|
|
8
8
|
|
|
9
9
|
Learn more in the [NanoGPT documentation](https://docs.nano-gpt.com).
|
|
10
10
|
|
|
@@ -221,6 +221,7 @@ for await (const chunk of stream) {
|
|
|
221
221
|
| `nano-gpt/glm-4.1v-thinking-flashx` | 64K | | | | | | $0.30 | $0.30 |
|
|
222
222
|
| `nano-gpt/GLM-4.6-Derestricted-v5` | 131K | | | | | | $0.40 | $2 |
|
|
223
223
|
| `nano-gpt/glm-z1-airx` | 32K | | | | | | $0.70 | $0.70 |
|
|
224
|
+
| `nano-gpt/google/diffusiongemma` | 262K | | | | | | $0.05 | $0.15 |
|
|
224
225
|
| `nano-gpt/google/gemini-3-flash-preview` | 1.0M | | | | | | $0.50 | $3 |
|
|
225
226
|
| `nano-gpt/google/gemini-3-flash-preview-thinking` | 1.0M | | | | | | $0.50 | $3 |
|
|
226
227
|
| `nano-gpt/google/gemini-3.1-flash-lite` | 1.0M | | | | | | $0.25 | $2 |
|
|
@@ -238,9 +239,11 @@ for await (const chunk of stream) {
|
|
|
238
239
|
| `nano-gpt/google/gemini-flash-lite-latest` | 1.0M | | | | | | $0.30 | $3 |
|
|
239
240
|
| `nano-gpt/google/gemini-pro-latest` | 1.0M | | | | | | $2 | $12 |
|
|
240
241
|
| `nano-gpt/google/gemma-4-26b-a4b-it` | 262K | | | | | | $0.12 | $0.38 |
|
|
242
|
+
| `nano-gpt/google/gemma-4-26b-a4b-it-cybersecurity` | 262K | | | | | | $0.11 | $0.33 |
|
|
241
243
|
| `nano-gpt/google/gemma-4-26b-a4b-it:thinking` | 262K | | | | | | $0.13 | $0.40 |
|
|
242
244
|
| `nano-gpt/google/gemma-4-31b-it` | 262K | | | | | | $0.10 | $0.45 |
|
|
243
245
|
| `nano-gpt/google/gemma-4-31b-it:thinking` | 262K | | | | | | $0.10 | $0.35 |
|
|
246
|
+
| `nano-gpt/google/gemma4-31b-splituntied` | 262K | | | | | | $0.10 | $0.30 |
|
|
244
247
|
| `nano-gpt/Gryphe/MythoMax-L2-13b` | 4K | | | | | | $0.10 | $0.10 |
|
|
245
248
|
| `nano-gpt/hermes-high` | 1.0M | | | | | | $1 | $3 |
|
|
246
249
|
| `nano-gpt/hermes-low` | 1.0M | | | | | | $1 | $3 |
|
|
@@ -350,6 +353,7 @@ for await (const chunk of stream) {
|
|
|
350
353
|
| `nano-gpt/nvidia/nemotron-3-super-120b-a12b:thinking` | 262K | | | | | | $0.05 | $0.25 |
|
|
351
354
|
| `nano-gpt/nvidia/nemotron-3-ultra-550b-a55b` | 1.0M | | | | | | $0.50 | $3 |
|
|
352
355
|
| `nano-gpt/nvidia/nemotron-3-ultra-550b-a55b:thinking` | 1.0M | | | | | | $0.50 | $3 |
|
|
356
|
+
| `nano-gpt/nvidia/nemotron-3.5-content-safety` | 131K | | | | | | $0.05 | $0.15 |
|
|
353
357
|
| `nano-gpt/nvidia/nemotron-3.5-lightning` | 1.0M | | | | | | $0.05 | $0.20 |
|
|
354
358
|
| `nano-gpt/nvidia/nemotron-3.5-lightning:thinking` | 1.0M | | | | | | $0.05 | $0.20 |
|
|
355
359
|
| `nano-gpt/openai/gpt-3.5-turbo` | 16K | | | | | | $0.50 | $2 |
|
|
@@ -467,12 +471,13 @@ for await (const chunk of stream) {
|
|
|
467
471
|
| `nano-gpt/qwen/qwen3.7-plus` | 992K | | | | | | $0.40 | $2 |
|
|
468
472
|
| `nano-gpt/qwen/qwen3.7-plus:thinking` | 984K | | | | | | $0.40 | $2 |
|
|
469
473
|
| `nano-gpt/qwen/qwen3.8-27b` | 262K | | | | | | $0.15 | $0.70 |
|
|
474
|
+
| `nano-gpt/qwen/qwen3.8-27b-cybersecurity` | 262K | | | | | | $0.10 | $0.60 |
|
|
470
475
|
| `nano-gpt/qwen/qwen3.8-27b-fable` | 524K | | | | | | $0.25 | $2 |
|
|
471
476
|
| `nano-gpt/qwen/qwen3.8-27b-obliterated` | 524K | | | | | | $0.25 | $2 |
|
|
472
477
|
| `nano-gpt/qwen/qwen3.8-27b-obliterated:thinking` | 524K | | | | | | $0.25 | $2 |
|
|
473
478
|
| `nano-gpt/qwen/qwen3.8-27b-queen` | 524K | | | | | | $0.25 | $2 |
|
|
474
|
-
| `nano-gpt/qwen/qwen3.8-27b-uncensored` | 524K | | | | | | $0.
|
|
475
|
-
| `nano-gpt/qwen/qwen3.8-27b-uncensored:thinking` | 524K | | | | | | $0.
|
|
479
|
+
| `nano-gpt/qwen/qwen3.8-27b-uncensored` | 524K | | | | | | $0.20 | $2 |
|
|
480
|
+
| `nano-gpt/qwen/qwen3.8-27b-uncensored:thinking` | 524K | | | | | | $0.20 | $2 |
|
|
476
481
|
| `nano-gpt/qwen/qwen3.8-27b:thinking` | 262K | | | | | | $0.15 | $0.70 |
|
|
477
482
|
| `nano-gpt/qwen/qwen3.8-flash` | 992K | | | | | | $0.14 | $0.42 |
|
|
478
483
|
| `nano-gpt/qwen/qwen3.8-max` | 991K | | | | | | $2 | $6 |
|
|
@@ -494,7 +499,6 @@ for await (const chunk of stream) {
|
|
|
494
499
|
| `nano-gpt/sarvam-105b` | 131K | | | | | | $0.05 | $0.21 |
|
|
495
500
|
| `nano-gpt/shisa-ai/shisa-v2-llama3.3-70b` | 128K | | | | | | $0.50 | $0.50 |
|
|
496
501
|
| `nano-gpt/shisa-ai/shisa-v2.1-llama3.3-70b` | 33K | | | | | | $0.50 | $0.50 |
|
|
497
|
-
| `nano-gpt/slowburn/gemma4-31b-splituntied` | 262K | | | | | | $0.10 | $0.30 |
|
|
498
502
|
| `nano-gpt/soob3123/amoral-gemma3-27B-v2` | 33K | | | | | | $0.30 | $0.30 |
|
|
499
503
|
| `nano-gpt/soob3123/GrayLine-Qwen3-8B` | 33K | | | | | | $0.30 | $0.30 |
|
|
500
504
|
| `nano-gpt/soob3123/Veiled-Calla-12B` | 33K | | | | | | $0.30 | $0.30 |
|
|
@@ -605,6 +609,7 @@ for await (const chunk of stream) {
|
|
|
605
609
|
| `nano-gpt/z-ai/glm-5.2:thinking` | 1.0M | | | | | | $0.42 | $1 |
|
|
606
610
|
| `nano-gpt/z-ai/glm-5.3` | 1.0M | | | | | | $1 | $3 |
|
|
607
611
|
| `nano-gpt/z-ai/glm-5.3-flash` | 1.0M | | | | | | $0.07 | $0.25 |
|
|
612
|
+
| `nano-gpt/z-ai/glm-5.3-flash-cybersecurity` | 1.0M | | | | | | $0.15 | $0.50 |
|
|
608
613
|
| `nano-gpt/z-ai/glm-5.3-flash-uncensored` | 1.0M | | | | | | $0.20 | $0.80 |
|
|
609
614
|
| `nano-gpt/z-ai/glm-5.3:thinking` | 1.0M | | | | | | $1 | $3 |
|
|
610
615
|
| `nano-gpt/z-ai/glm-5v-turbo` | 203K | | | | | | $1 | $4 |
|
|
@@ -4,7 +4,7 @@
|
|
|
4
4
|
|
|
5
5
|
# Nebius Token Factory
|
|
6
6
|
|
|
7
|
-
Access
|
|
7
|
+
Access 20 Nebius Token Factory models through Mastra's model router. Authentication is handled automatically using the `NEBIUS_API_KEY` environment variable.
|
|
8
8
|
|
|
9
9
|
Learn more in the [Nebius Token Factory documentation](https://docs.tokenfactory.nebius.com/).
|
|
10
10
|
|
|
@@ -40,6 +40,8 @@ for await (const chunk of stream) {
|
|
|
40
40
|
| ------------------------------------------- | ------- | ----- | --------- | ----- | ----- | ----- | ---------- | ----------- |
|
|
41
41
|
| `nebius/deepseek-ai/DeepSeek-V4-Flash-0731` | 1.0M | | | | | | $0.14 | $0.28 |
|
|
42
42
|
| `nebius/deepseek-ai/DeepSeek-V4-Pro` | 1.0M | | | | | | $2 | $4 |
|
|
43
|
+
| `nebius/deepseek-ai/DeepSeek-V4-Pro-0813` | 979K | | | | | | $1 | $4 |
|
|
44
|
+
| `nebius/deepseek-ai/DeepSeek-V4.1-Flash` | 1.0M | | | | | | $0.30 | $1 |
|
|
43
45
|
| `nebius/google/gemma-3-27b-it` | 110K | | | | | | $0.10 | $0.30 |
|
|
44
46
|
| `nebius/MiniMaxAI/MiniMax-M3` | 1.0M | | | | | | $0.30 | $1 |
|
|
45
47
|
| `nebius/moonshotai/Kimi-K2.7-Code` | 262K | | | | | | $0.95 | $4 |
|
|
@@ -54,6 +56,7 @@ for await (const chunk of stream) {
|
|
|
54
56
|
| `nebius/Qwen/Qwen3-Embedding-8B` | 41K | | | | | | $0.01 | — |
|
|
55
57
|
| `nebius/Qwen/Qwen3.5-397B-A17B` | 262K | | | | | | $0.60 | $4 |
|
|
56
58
|
| `nebius/zai-org/GLM-5.2` | 1.0M | | | | | | $1 | $4 |
|
|
59
|
+
| `nebius/zai-org/GLM-5.3` | 1.0M | | | | | | $1 | $4 |
|
|
57
60
|
| `nebius/zai-org/GLM-5.3-Flash` | 1.0M | | | | | | $0.15 | $0.50 |
|
|
58
61
|
|
|
59
62
|
Model availability, capabilities, context windows, and pricing are sourced from [models.dev](https://models.dev) and may change.
|
|
@@ -4,7 +4,7 @@
|
|
|
4
4
|
|
|
5
5
|
# OpenCode Zen
|
|
6
6
|
|
|
7
|
-
Access
|
|
7
|
+
Access 107 OpenCode Zen models through Mastra's model router. Authentication is handled automatically using the `OPENCODE_API_KEY` environment variable.
|
|
8
8
|
|
|
9
9
|
Learn more in the [OpenCode Zen documentation](https://opencode.ai/docs/zen).
|
|
10
10
|
|
|
@@ -54,6 +54,7 @@ for await (const chunk of stream) {
|
|
|
54
54
|
| `opencode/deepseek-v4-flash` | 1.0M | | | | | | $0.14 | $0.28 |
|
|
55
55
|
| `opencode/deepseek-v4-flash-vision-exp` | 1.0M | | | | | | $0.14 | $0.28 |
|
|
56
56
|
| `opencode/deepseek-v4-pro` | 1.0M | | | | | | $2 | $4 |
|
|
57
|
+
| `opencode/deepseek-v4.1-flash` | 1.0M | | | | | | $0.30 | $1 |
|
|
57
58
|
| `opencode/gemini-3-flash` | 1.0M | | | | | | $0.50 | $3 |
|
|
58
59
|
| `opencode/gemini-3.1-pro` | 1.0M | | | | | | $2 | $12 |
|
|
59
60
|
| `opencode/gemini-3.5-flash` | 1.0M | | | | | | $2 | $9 |
|
|
@@ -90,6 +91,8 @@ for await (const chunk of stream) {
|
|
|
90
91
|
| `opencode/grok-4.5` | 500K | | | | | | $2 | $6 |
|
|
91
92
|
| `opencode/grok-4.6` | 500K | | | | | | $2 | $6 |
|
|
92
93
|
| `opencode/grok-build-0.1` | 256K | | | | | | $1 | $2 |
|
|
94
|
+
| `opencode/jev-1.13` | 64K | | | | | | $0.04 | — |
|
|
95
|
+
| `opencode/jev-1.13-free` | 64K | | | | | | — | — |
|
|
93
96
|
| `opencode/jev-latest` | 64K | | | | | | $0.04 | — |
|
|
94
97
|
| `opencode/kimi-k2.5` | 262K | | | | | | $0.60 | $3 |
|
|
95
98
|
| `opencode/kimi-k2.6` | 262K | | | | | | $0.95 | $4 |
|
|
@@ -108,6 +111,7 @@ for await (const chunk of stream) {
|
|
|
108
111
|
| `opencode/nemotron-3.5-lightning-free` | 262K | | | | | | — | — |
|
|
109
112
|
| `opencode/qwen3.5-plus` | 262K | | | | | | $0.20 | $1 |
|
|
110
113
|
| `opencode/qwen3.6-plus` | 262K | | | | | | $0.50 | $3 |
|
|
114
|
+
| `opencode/qwen3.8-flash` | 1.0M | | | | | | $0.15 | $0.47 |
|
|
111
115
|
|
|
112
116
|
Model availability, capabilities, context windows, and pricing are sourced from [models.dev](https://models.dev) and may change.
|
|
113
117
|
|
|
@@ -4,7 +4,7 @@
|
|
|
4
4
|
|
|
5
5
|
# TensorX
|
|
6
6
|
|
|
7
|
-
Access
|
|
7
|
+
Access 25 TensorX models through Mastra's model router. Authentication is handled automatically using the `TENSORX_API_KEY` environment variable.
|
|
8
8
|
|
|
9
9
|
Learn more in the [TensorX documentation](https://docs.tensorx.ai/).
|
|
10
10
|
|
|
@@ -19,7 +19,7 @@ const agent = new Agent({
|
|
|
19
19
|
id: "my-agent",
|
|
20
20
|
name: "My Agent",
|
|
21
21
|
instructions: "You are a helpful assistant",
|
|
22
|
-
model: "tensorx/deepseek/deepseek-
|
|
22
|
+
model: "tensorx/deepseek/deepseek-r1-0528"
|
|
23
23
|
});
|
|
24
24
|
|
|
25
25
|
// Generate a response
|
|
@@ -36,34 +36,33 @@ for await (const chunk of stream) {
|
|
|
36
36
|
|
|
37
37
|
## Models
|
|
38
38
|
|
|
39
|
-
| Model
|
|
40
|
-
|
|
|
41
|
-
| `tensorx/deepseek/deepseek-
|
|
42
|
-
| `tensorx/deepseek/deepseek-
|
|
43
|
-
| `tensorx/deepseek/deepseek-
|
|
44
|
-
| `tensorx/deepseek/deepseek-v4-
|
|
45
|
-
| `tensorx/deepseek/deepseek-v4-
|
|
46
|
-
| `tensorx/deepseek/deepseek-v4-
|
|
47
|
-
| `tensorx/
|
|
48
|
-
| `tensorx/minimax/minimax-
|
|
49
|
-
| `tensorx/
|
|
50
|
-
| `tensorx/moonshotai/kimi-k2.
|
|
51
|
-
| `tensorx/moonshotai/kimi-k2.
|
|
52
|
-
| `tensorx/moonshotai/kimi-
|
|
53
|
-
| `tensorx/
|
|
54
|
-
| `tensorx/
|
|
55
|
-
| `tensorx/
|
|
56
|
-
| `tensorx/qwen/qwen3-
|
|
57
|
-
| `tensorx/qwen/qwen3-
|
|
58
|
-
| `tensorx/qwen/qwen3-
|
|
59
|
-
| `tensorx/
|
|
60
|
-
| `tensorx/
|
|
61
|
-
| `tensorx/z-ai/glm-
|
|
62
|
-
| `tensorx/z-ai/glm-5`
|
|
63
|
-
| `tensorx/z-ai/glm-5
|
|
64
|
-
| `tensorx/z-ai/glm-5.
|
|
65
|
-
| `tensorx/z-ai/glm-
|
|
66
|
-
| `tensorx/z-ai/glm-5v-turbo` | 203K | | | | | | $1 | $4 |
|
|
39
|
+
| Model | Context | Tools | Reasoning | Image | Audio | Video | Input $/1M | Output $/1M |
|
|
40
|
+
| ----------------------------------------- | ------- | ----- | --------- | ----- | ----- | ----- | ---------- | ----------- |
|
|
41
|
+
| `tensorx/deepseek/deepseek-r1-0528` | 164K | | | | | | $0.66 | $3 |
|
|
42
|
+
| `tensorx/deepseek/deepseek-v3.2` | 164K | | | | | | $0.30 | $0.50 |
|
|
43
|
+
| `tensorx/deepseek/deepseek-v4-flash-0731` | 1.0M | | | | | | $0.25 | $0.30 |
|
|
44
|
+
| `tensorx/deepseek/deepseek-v4-pro` | 1.0M | | | | | | $2 | $4 |
|
|
45
|
+
| `tensorx/deepseek/deepseek-v4-pro-0813` | 1.0M | | | | | | $2 | $4 |
|
|
46
|
+
| `tensorx/deepseek/deepseek-v4.1-flash` | 1.0M | | | | | | $0.50 | $2 |
|
|
47
|
+
| `tensorx/minimax/minimax-m2.5` | 197K | | | | | | $0.30 | $1 |
|
|
48
|
+
| `tensorx/minimax/minimax-m3` | 1.0M | | | | | | $0.40 | $2 |
|
|
49
|
+
| `tensorx/moonshotai/kimi-k2.5` | 262K | | | | | | $0.50 | $3 |
|
|
50
|
+
| `tensorx/moonshotai/kimi-k2.6` | 262K | | | | | | $1 | $4 |
|
|
51
|
+
| `tensorx/moonshotai/kimi-k2.7-code` | 262K | | | | | | $1 | $5 |
|
|
52
|
+
| `tensorx/moonshotai/kimi-k3` | 1.0M | | | | | | $3 | $15 |
|
|
53
|
+
| `tensorx/qwen/qwen3-235b-a22b-2507` | 131K | | | | | | $0.07 | $0.46 |
|
|
54
|
+
| `tensorx/qwen/qwen3.5-122b-a10b` | 262K | | | | | | $0.50 | $4 |
|
|
55
|
+
| `tensorx/qwen/qwen3.5-9b` | 262K | | | | | | $0.15 | $0.20 |
|
|
56
|
+
| `tensorx/qwen/qwen3.8-2.4t-a95b` | 262K | | | | | | $3 | $6 |
|
|
57
|
+
| `tensorx/qwen/qwen3.8-27b` | 262K | | | | | | $0.40 | $2 |
|
|
58
|
+
| `tensorx/qwen/qwen3.8-flash-next` | 262K | | | | | | $0.20 | $0.50 |
|
|
59
|
+
| `tensorx/z-ai/glm-5` | 203K | | | | | | $1 | $3 |
|
|
60
|
+
| `tensorx/z-ai/glm-5-turbo` | 203K | | | | | | $1 | $4 |
|
|
61
|
+
| `tensorx/z-ai/glm-5.1` | 203K | | | | | | $1 | $4 |
|
|
62
|
+
| `tensorx/z-ai/glm-5.2` | 1.0M | | | | | | $2 | $5 |
|
|
63
|
+
| `tensorx/z-ai/glm-5.3` | 1.0M | | | | | | $2 | $5 |
|
|
64
|
+
| `tensorx/z-ai/glm-5.3-flash` | 1.0M | | | | | | $0.20 | $0.50 |
|
|
65
|
+
| `tensorx/z-ai/glm-5v-turbo` | 203K | | | | | | $1 | $4 |
|
|
67
66
|
|
|
68
67
|
Model availability, capabilities, context windows, and pricing are sourced from [models.dev](https://models.dev) and may change.
|
|
69
68
|
|
|
@@ -77,7 +76,7 @@ const agent = new Agent({
|
|
|
77
76
|
name: "custom-agent",
|
|
78
77
|
model: {
|
|
79
78
|
url: "https://api.tensorx.ai/v1",
|
|
80
|
-
id: "tensorx/deepseek/deepseek-
|
|
79
|
+
id: "tensorx/deepseek/deepseek-r1-0528",
|
|
81
80
|
apiKey: process.env.TENSORX_API_KEY,
|
|
82
81
|
headers: {
|
|
83
82
|
"X-Custom-Header": "value"
|
|
@@ -96,7 +95,7 @@ const agent = new Agent({
|
|
|
96
95
|
const useAdvanced = requestContext.task === "complex";
|
|
97
96
|
return useAdvanced
|
|
98
97
|
? "tensorx/z-ai/glm-5v-turbo"
|
|
99
|
-
: "tensorx/deepseek/deepseek-
|
|
98
|
+
: "tensorx/deepseek/deepseek-r1-0528";
|
|
100
99
|
}
|
|
101
100
|
});
|
|
102
101
|
```
|
|
@@ -4,7 +4,7 @@
|
|
|
4
4
|
|
|
5
5
|
# Z.AI
|
|
6
6
|
|
|
7
|
-
Access
|
|
7
|
+
Access 17 Z.AI models through Mastra's model router. Authentication is handled automatically using the `ZHIPU_API_KEY` environment variable.
|
|
8
8
|
|
|
9
9
|
Learn more in the [Z.AI documentation](https://docs.z.ai/guides/overview/pricing).
|
|
10
10
|
|
|
@@ -52,7 +52,8 @@ for await (const chunk of stream) {
|
|
|
52
52
|
| `zai/glm-5.1` | 200K | | | | | | $1 | $4 |
|
|
53
53
|
| `zai/glm-5.2` | 1.0M | | | | | | $1 | $4 |
|
|
54
54
|
| `zai/glm-5.3` | 1.0M | | | | | | $1 | $4 |
|
|
55
|
-
| `zai/glm-5.3-flash` | 1.0M | | | | | | $0.
|
|
55
|
+
| `zai/glm-5.3-flash` | 1.0M | | | | | | $0.15 | $0.50 |
|
|
56
|
+
| `zai/glm-5.3-flashx` | 1.0M | | | | | | $0.37 | $1 |
|
|
56
57
|
| `zai/glm-5v-turbo` | 200K | | | | | | $1 | $4 |
|
|
57
58
|
|
|
58
59
|
Model availability, capabilities, context windows, and pricing are sourced from [models.dev](https://models.dev) and may change.
|
|
@@ -4,7 +4,7 @@
|
|
|
4
4
|
|
|
5
5
|
# ZenMux
|
|
6
6
|
|
|
7
|
-
Access
|
|
7
|
+
Access 122 ZenMux models through Mastra's model router. Authentication is handled automatically using the `ZENMUX_API_KEY` environment variable.
|
|
8
8
|
|
|
9
9
|
Learn more in the [ZenMux documentation](https://docs.zenmux.ai).
|
|
10
10
|
|
|
@@ -117,31 +117,38 @@ for await (const chunk of stream) {
|
|
|
117
117
|
| `zenmux/volcengine/doubao-seed-2.0-lite` | 256K | | | | | | $0.09 | $0.51 |
|
|
118
118
|
| `zenmux/volcengine/doubao-seed-2.0-mini` | 256K | | | | | | $0.03 | $0.28 |
|
|
119
119
|
| `zenmux/volcengine/doubao-seed-2.0-pro` | 256K | | | | | | $0.45 | $2 |
|
|
120
|
-
| `zenmux/x-ai/grok-4.2-fast` | 2.0M | | | | | | $
|
|
121
|
-
| `zenmux/x-ai/grok-4.2-fast-non-reasoning` | 2.0M | | | | | | $
|
|
120
|
+
| `zenmux/x-ai/grok-4.2-fast` | 2.0M | | | | | | $2 | $6 |
|
|
121
|
+
| `zenmux/x-ai/grok-4.2-fast-non-reasoning` | 2.0M | | | | | | $2 | $6 |
|
|
122
122
|
| `zenmux/x-ai/grok-4.3` | 1.0M | | | | | | $1 | $3 |
|
|
123
123
|
| `zenmux/x-ai/grok-4.5` | 500K | | | | | | $2 | $6 |
|
|
124
|
+
| `zenmux/x-ai/grok-4.6` | 500K | | | | | | $2 | $6 |
|
|
124
125
|
| `zenmux/x-ai/grok-build-0.1` | 256K | | | | | | $1 | $2 |
|
|
126
|
+
| `zenmux/x-ai/grok-imagine-image-2.0` | 66K | | | | | | — | — |
|
|
127
|
+
| `zenmux/x-ai/grok-voice-stt-1.0` | 15K | | | | | | — | — |
|
|
128
|
+
| `zenmux/x-ai/grok-voice-tts-1.0` | 15K | | | | | | — | — |
|
|
125
129
|
| `zenmux/xiaomi/mimo-v2-flash` | 262K | | | | | | $0.10 | $0.30 |
|
|
126
130
|
| `zenmux/xiaomi/mimo-v2-omni` | 265K | | | | | | $0.40 | $2 |
|
|
127
131
|
| `zenmux/xiaomi/mimo-v2-pro` | 1.0M | | | | | | $1 | $3 |
|
|
128
132
|
| `zenmux/xiaomi/mimo-v2.5` | 1.0M | | | | | | $0.40 | $2 |
|
|
129
133
|
| `zenmux/xiaomi/mimo-v2.5-pro` | 1.0M | | | | | | $1 | $3 |
|
|
130
|
-
| `zenmux/z-ai/glm-4.5` | 128K | | | | | | $0.
|
|
131
|
-
| `zenmux/z-ai/glm-4.5-air` | 128K | | | | | | $0.
|
|
132
|
-
| `zenmux/z-ai/glm-4.6` | 200K | | | | | | $0.
|
|
133
|
-
| `zenmux/z-ai/glm-4.6v` | 200K | | | | | | $0.
|
|
134
|
-
| `zenmux/z-ai/glm-4.6v-flash` | 200K | | | | | | $0.02 | $0.
|
|
134
|
+
| `zenmux/z-ai/glm-4.5` | 128K | | | | | | $0.29 | $1 |
|
|
135
|
+
| `zenmux/z-ai/glm-4.5-air` | 128K | | | | | | $0.12 | $0.29 |
|
|
136
|
+
| `zenmux/z-ai/glm-4.6` | 200K | | | | | | $0.29 | $1 |
|
|
137
|
+
| `zenmux/z-ai/glm-4.6v` | 200K | | | | | | $0.15 | $0.44 |
|
|
138
|
+
| `zenmux/z-ai/glm-4.6v-flash` | 200K | | | | | | $0.02 | $0.22 |
|
|
135
139
|
| `zenmux/z-ai/glm-4.6v-flash-free` | 200K | | | | | | — | — |
|
|
136
|
-
| `zenmux/z-ai/glm-4.7` | 200K | | | | | | $0.
|
|
140
|
+
| `zenmux/z-ai/glm-4.7` | 200K | | | | | | $0.29 | $1 |
|
|
137
141
|
| `zenmux/z-ai/glm-4.7-flash-free` | 200K | | | | | | — | — |
|
|
138
|
-
| `zenmux/z-ai/glm-4.7-flashx` | 200K | | | | | | $0.07 | $0.
|
|
142
|
+
| `zenmux/z-ai/glm-4.7-flashx` | 200K | | | | | | $0.07 | $0.44 |
|
|
139
143
|
| `zenmux/z-ai/glm-5` | 200K | | | | | | $0.58 | $3 |
|
|
140
|
-
| `zenmux/z-ai/glm-5-turbo` | 200K | | | | | | $0.
|
|
144
|
+
| `zenmux/z-ai/glm-5-turbo` | 200K | | | | | | $0.73 | $3 |
|
|
141
145
|
| `zenmux/z-ai/glm-5.1` | 200K | | | | | | $0.88 | $4 |
|
|
142
|
-
| `zenmux/z-ai/glm-5.2` | 1.0M | | | | | | $
|
|
143
|
-
| `zenmux/z-ai/glm-5.
|
|
146
|
+
| `zenmux/z-ai/glm-5.2` | 1.0M | | | | | | $0.98 | $3 |
|
|
147
|
+
| `zenmux/z-ai/glm-5.3` | 1.0M | | | | | | $1 | $4 |
|
|
148
|
+
| `zenmux/z-ai/glm-5.3-flash` | 1.0M | | | | | | $0.15 | $0.50 |
|
|
149
|
+
| `zenmux/z-ai/glm-5.3-flashx` | 1.0M | | | | | | $0.38 | $1 |
|
|
144
150
|
| `zenmux/z-ai/glm-5v-turbo` | 200K | | | | | | $0.73 | $3 |
|
|
151
|
+
| `zenmux/z-ai/glm-image` | 10K | | | | | | — | — |
|
|
145
152
|
|
|
146
153
|
Model availability, capabilities, context windows, and pricing are sourced from [models.dev](https://models.dev) and may change.
|
|
147
154
|
|
|
@@ -173,7 +180,7 @@ const agent = new Agent({
|
|
|
173
180
|
model: ({ requestContext }) => {
|
|
174
181
|
const useAdvanced = requestContext.task === "complex";
|
|
175
182
|
return useAdvanced
|
|
176
|
-
? "zenmux/z-ai/glm-
|
|
183
|
+
? "zenmux/z-ai/glm-image"
|
|
177
184
|
: "zenmux/anthropic/claude-3.5-haiku";
|
|
178
185
|
}
|
|
179
186
|
});
|
|
@@ -4,7 +4,7 @@
|
|
|
4
4
|
|
|
5
5
|
# Zhipu AI
|
|
6
6
|
|
|
7
|
-
Access
|
|
7
|
+
Access 16 Zhipu AI models through Mastra's model router. Authentication is handled automatically using the `ZHIPU_API_KEY` environment variable.
|
|
8
8
|
|
|
9
9
|
Learn more in the [Zhipu AI documentation](https://docs.z.ai/guides/overview/pricing).
|
|
10
10
|
|
|
@@ -51,7 +51,8 @@ for await (const chunk of stream) {
|
|
|
51
51
|
| `zhipuai/glm-5.1` | 200K | | | | | | $1 | $4 |
|
|
52
52
|
| `zhipuai/glm-5.2` | 1.0M | | | | | | $1 | $4 |
|
|
53
53
|
| `zhipuai/glm-5.3` | 1.0M | | | | | | $1 | $4 |
|
|
54
|
-
| `zhipuai/glm-5.3-flash` | 1.0M | | | | | | $0.
|
|
54
|
+
| `zhipuai/glm-5.3-flash` | 1.0M | | | | | | $0.15 | $0.50 |
|
|
55
|
+
| `zhipuai/glm-5.3-flashx` | 1.0M | | | | | | $0.37 | $1 |
|
|
55
56
|
| `zhipuai/glm-5v-turbo` | 200K | | | | | | $5 | $22 |
|
|
56
57
|
|
|
57
58
|
Model availability, capabilities, context windows, and pricing are sourced from [models.dev](https://models.dev) and may change.
|
|
@@ -111,10 +111,10 @@ Direct access to individual AI model providers. Each provider offers unique mode
|
|
|
111
111
|
- [Meganova](https://mastra.ai/models/providers/meganova)
|
|
112
112
|
- [Melious](https://mastra.ai/models/providers/melious)
|
|
113
113
|
- [Meta](https://mastra.ai/models/providers/meta)
|
|
114
|
+
- [MiniMax (minimax.cn)](https://mastra.ai/models/providers/minimax-cn)
|
|
114
115
|
- [MiniMax (minimax.io)](https://mastra.ai/models/providers/minimax)
|
|
115
|
-
- [MiniMax (
|
|
116
|
+
- [MiniMax Token Plan (minimax.cn)](https://mastra.ai/models/providers/minimax-cn-coding-plan)
|
|
116
117
|
- [MiniMax Token Plan (minimax.io)](https://mastra.ai/models/providers/minimax-coding-plan)
|
|
117
|
-
- [MiniMax Token Plan (minimaxi.com)](https://mastra.ai/models/providers/minimax-cn-coding-plan)
|
|
118
118
|
- [Mixlayer](https://mastra.ai/models/providers/mixlayer)
|
|
119
119
|
- [Moark](https://mastra.ai/models/providers/moark)
|
|
120
120
|
- [Modal](https://mastra.ai/models/providers/modal)
|
|
@@ -168,7 +168,7 @@ const result = await agent.generate('message for agent')
|
|
|
168
168
|
|
|
169
169
|
**options.modelSettings.frequencyPenalty** (`number`): Penalty for token frequency (-2 to 2). Reduces repetition of frequent tokens.
|
|
170
170
|
|
|
171
|
-
**options.modelSettings.timeout** (`object`): Time-based execution budget for the run. Accepts totalMs, the maximum duration of the entire agent run across every loop iteration, tool call and retry, and stepMs, the maximum duration of a single model call including the time spent consuming its stream. Exceeding either budget fails with a MastraTimeoutError. A totalMs timeout ends the run and does not try fallback models, because it is a hard deadline for the whole run. A stepMs timeout is not retried against the same model but does advance to the next entry in models when fallback models are configured. Also accepts firstChunkMs, which only applies to streaming calls and is the maximum time the model may take to emit its first content-bearing chunk (text, reasoning, tool call, file or source; stream-start and metadata chunks do not count). A firstChunkMs timeout fails with timeoutType: 'firstChunk', behaves like stepMs for fallback, and is reset for each provider retry attempt. Nested timeout keys are merged across call-time and per-model settings.
|
|
171
|
+
**options.modelSettings.timeout** (`object`): Time-based execution budget for the run. It must be an object whose configured values are positive, finite numbers of milliseconds. Accepts totalMs, the maximum duration of the entire agent run across every loop iteration, tool call and retry, and stepMs, the maximum duration of a single model call including the time spent consuming its stream. Exceeding either budget fails with a MastraTimeoutError. A totalMs timeout ends the run and does not try fallback models, because it is a hard deadline for the whole run. A stepMs timeout is not retried against the same model but does advance to the next entry in models when fallback models are configured. Also accepts firstChunkMs, which only applies to streaming calls and is the maximum time the model may take to emit its first content-bearing chunk (text, reasoning, tool call, file or source; stream-start and metadata chunks do not count). A firstChunkMs timeout fails with timeoutType: 'firstChunk', behaves like stepMs for fallback, and is reset for each provider retry attempt. Nested timeout keys are merged across call-time and per-model settings.
|
|
172
172
|
|
|
173
173
|
**options.modelSettings.stopSequences** (`string[]`): Stop sequences. If set, the model will stop generating text when one of the stop sequences is generated.
|
|
174
174
|
|