@mastra/mcp-docs-server 1.2.27-alpha.15 → 1.2.27-alpha.19

This diff represents the content of publicly available package versions that have been released to one of the supported registries. The information contained in this diff is provided for informational purposes only and reflects changes between package versions as they appear in their respective public registries.
@@ -4,7 +4,7 @@
4
4
 
5
5
  # ![OpenRouter logo](https://models.dev/logos/openrouter.svg)OpenRouter
6
6
 
7
- OpenRouter aggregates models from multiple providers with enhanced features like rate limiting and failover. Access 371 models through Mastra's model router.
7
+ OpenRouter aggregates models from multiple providers with enhanced features like rate limiting and failover. Access 372 models through Mastra's model router.
8
8
 
9
9
  Learn more in the [OpenRouter documentation](https://openrouter.ai/models).
10
10
 
@@ -408,4 +408,5 @@ ANTHROPIC_API_KEY=ant-...
408
408
  | `z-ai/glm-5.2:free` |
409
409
  | `z-ai/glm-5.3` |
410
410
  | `z-ai/glm-5.3-flash` |
411
+ | `z-ai/glm-5.3-flashx` |
411
412
  | `z-ai/glm-5v-turbo` |
@@ -4,7 +4,7 @@
4
4
 
5
5
  # ![Vercel logo](https://models.dev/logos/vercel.svg)Vercel
6
6
 
7
- Vercel aggregates models from multiple providers with enhanced features like rate limiting and failover. Access 372 models through Mastra's model router.
7
+ Vercel aggregates models from multiple providers with enhanced features like rate limiting and failover. Access 375 models through Mastra's model router.
8
8
 
9
9
  Learn more in the [Vercel documentation](https://ai-sdk.dev/providers/ai-sdk-providers).
10
10
 
@@ -234,6 +234,7 @@ ANTHROPIC_API_KEY=ant-...
234
234
  | `mistral/mistral-medium-3.5` |
235
235
  | `mistral/mistral-nemo` |
236
236
  | `mistral/mistral-small` |
237
+ | `mixedbread/toast-1` |
237
238
  | `moonshotai/kimi-k2` |
238
239
  | `moonshotai/kimi-k2-thinking` |
239
240
  | `moonshotai/kimi-k2.5` |
@@ -337,6 +338,8 @@ ANTHROPIC_API_KEY=ant-...
337
338
  | `poolside/laguna-s-2.1-free` |
338
339
  | `prodia/flux-fast-schnell` |
339
340
  | `quiverai/arrow-1.1` |
341
+ | `quiverai/arrow-2` |
342
+ | `quiverai/arrow-2-telos` |
340
343
  | `recraft/recraft-v2` |
341
344
  | `recraft/recraft-v3` |
342
345
  | `recraft/recraft-v4` |
@@ -4,7 +4,7 @@
4
4
 
5
5
  # Model Providers
6
6
 
7
- Mastra provides a unified interface for working with LLMs across multiple providers, giving you access to 7352 models from 209 providers through a single API.
7
+ Mastra provides a unified interface for working with LLMs across multiple providers, giving you access to 7369 models from 209 providers through a single API.
8
8
 
9
9
  ## Features
10
10
 
@@ -42,12 +42,12 @@ for await (const chunk of stream) {
42
42
  | `kilo/~anthropic/claude-haiku-latest` | 200K | | | | | | $1 | $5 |
43
43
  | `kilo/~anthropic/claude-opus-latest` | 1.0M | | | | | | $5 | $25 |
44
44
  | `kilo/~anthropic/claude-sonnet-latest` | 1.0M | | | | | | $2 | $10 |
45
- | `kilo/~deepseek/deepseek-flash-latest` | 1.0M | | | | | | $0.14 | $0.54 |
45
+ | `kilo/~deepseek/deepseek-flash-latest` | 1.0M | | | | | | $0.13 | $0.52 |
46
46
  | `kilo/~deepseek/deepseek-pro-latest` | 1.0M | | | | | | $0.58 | $2 |
47
- | `kilo/~deepseek/deepseek-v4-flash-latest` | 1.0M | | | | | | $0.05 | $0.16 |
47
+ | `kilo/~deepseek/deepseek-v4-flash-latest` | 1.0M | | | | | | $0.04 | $0.08 |
48
48
  | `kilo/~google/gemini-flash-latest` | 1.0M | | | | | | $0.75 | $4 |
49
49
  | `kilo/~google/gemini-pro-latest` | 1.0M | | | | | | $2 | $12 |
50
- | `kilo/~moonshotai/kimi-latest` | 1.0M | | | | | | $2 | $11 |
50
+ | `kilo/~moonshotai/kimi-latest` | 1.0M | | | | | | $2 | $9 |
51
51
  | `kilo/~openai/gpt-astra-latest` | 1.1M | | | | | | $10 | $50 |
52
52
  | `kilo/~openai/gpt-luna-latest` | 1.1M | | | | | | $0.20 | $1 |
53
53
  | `kilo/~openai/gpt-mini-latest` | 400K | | | | | | $0.75 | $5 |
@@ -55,7 +55,7 @@ for await (const chunk of stream) {
55
55
  | `kilo/~openai/gpt-terra-latest` | 1.1M | | | | | | $2 | $12 |
56
56
  | `kilo/~x-ai/grok-latest` | 500K | | | | | | $2 | $6 |
57
57
  | `kilo/~z-ai/glm-flash-latest` | 1.0M | | | | | | $0.07 | $0.25 |
58
- | `kilo/~z-ai/glm-latest` | 1.0M | | | | | | $0.89 | $3 |
58
+ | `kilo/~z-ai/glm-latest` | 1.0M | | | | | | $0.84 | $3 |
59
59
  | `kilo/aion-labs/aion-2.0` | 131K | | | | | | $0.80 | $2 |
60
60
  | `kilo/aion-labs/aion-3.0` | 131K | | | | | | $3 | $6 |
61
61
  | `kilo/aion-labs/aion-3.0-mini` | 131K | | | | | | $0.70 | $1 |
@@ -212,7 +212,7 @@ for await (const chunk of stream) {
212
212
  | `kilo/moonshotai/kimi-k2.5` | 262K | | | | | | $0.60 | $3 |
213
213
  | `kilo/moonshotai/kimi-k2.6` | 262K | | | | | | $0.80 | $3 |
214
214
  | `kilo/moonshotai/kimi-k2.7-code` | 262K | | | | | | $0.95 | $4 |
215
- | `kilo/moonshotai/kimi-k3` | 1.0M | | | | | | $2 | $11 |
215
+ | `kilo/moonshotai/kimi-k3` | 1.0M | | | | | | $2 | $9 |
216
216
  | `kilo/morph/morph-v3-fast` | 82K | | | | | | $0.80 | $1 |
217
217
  | `kilo/morph/morph-v3-large` | 262K | | | | | | $0.90 | $2 |
218
218
  | `kilo/nex-agi/nex-n2.5-mini:free` | 262K | | | | | | — | — |
@@ -224,11 +224,11 @@ for await (const chunk of stream) {
224
224
  | `kilo/nvidia/nemotron-3-nano-omni-30b-a3b-reasoning:free` | 256K | | | | | | — | — |
225
225
  | `kilo/nvidia/nemotron-3-super-120b-a12b` | 262K | | | | | | $0.08 | $0.45 |
226
226
  | `kilo/nvidia/nemotron-3-super-120b-a12b:free` | 262K | | | | | | — | — |
227
- | `kilo/nvidia/nemotron-3-ultra-550b-a55b` | 256K | | | | | | $0.50 | $2 |
227
+ | `kilo/nvidia/nemotron-3-ultra-550b-a55b` | 203K | | | | | | $0.50 | $2 |
228
228
  | `kilo/nvidia/nemotron-3-ultra-550b-a55b:free` | 1.0M | | | | | | — | — |
229
229
  | `kilo/nvidia/nemotron-3.5-content-safety` | 131K | | | | | | $0.20 | $0.20 |
230
230
  | `kilo/nvidia/nemotron-3.5-content-safety:free` | 128K | | | | | | — | — |
231
- | `kilo/nvidia/nemotron-3.5-lightning` | 262K | | | | | | $0.07 | $0.18 |
231
+ | `kilo/nvidia/nemotron-3.5-lightning` | 262K | | | | | | $0.04 | $0.18 |
232
232
  | `kilo/nvidia/nemotron-3.5-lightning:free` | 1.0M | | | | | | — | — |
233
233
  | `kilo/openai/gpt-3.5-turbo` | 16K | | | | | | $0.50 | $2 |
234
234
  | `kilo/openai/gpt-3.5-turbo-0613` | 4K | | | | | | $1 | $2 |
@@ -270,7 +270,6 @@ for await (const chunk of stream) {
270
270
  | `kilo/openai/gpt-5.6-luna` | 1.1M | | | | | | $0.20 | $1 |
271
271
  | `kilo/openai/gpt-5.6-luna-pro` | 1.1M | | | | | | $0.20 | $1 |
272
272
  | `kilo/openai/gpt-5.6-sol` | 1.1M | | | | | | $4 | $20 |
273
- | `kilo/openai/gpt-5.6-sol-discounted` | 1.1M | | | | | | $2 | $10 |
274
273
  | `kilo/openai/gpt-5.6-sol-pro` | 1.1M | | | | | | $4 | $20 |
275
274
  | `kilo/openai/gpt-5.6-terra` | 1.1M | | | | | | $2 | $12 |
276
275
  | `kilo/openai/gpt-5.6-terra-pro` | 1.1M | | | | | | $2 | $12 |
@@ -280,7 +279,7 @@ for await (const chunk of stream) {
280
279
  | `kilo/openai/gpt-audio-mini` | 128K | | | | | | $0.60 | $2 |
281
280
  | `kilo/openai/gpt-chat-latest` | 400K | | | | | | $5 | $30 |
282
281
  | `kilo/openai/gpt-oss-120b` | 131K | | | | | | $0.03 | $0.17 |
283
- | `kilo/openai/gpt-oss-20b` | 131K | | | | | | $0.02 | $0.10 |
282
+ | `kilo/openai/gpt-oss-20b` | 131K | | | | | | $0.02 | $0.09 |
284
283
  | `kilo/openai/gpt-oss-safeguard-20b` | 131K | | | | | | $0.07 | $0.30 |
285
284
  | `kilo/openai/o1` | 200K | | | | | | $15 | $60 |
286
285
  | `kilo/openai/o1-pro` | 200K | | | | | | $150 | $600 |
@@ -416,6 +415,7 @@ for await (const chunk of stream) {
416
415
  | `kilo/z-ai/glm-5.2:free` | 33K | | | | | | — | — |
417
416
  | `kilo/z-ai/glm-5.3` | 1.0M | | | | | | $1 | $4 |
418
417
  | `kilo/z-ai/glm-5.3-flash` | 1.0M | | | | | | $0.15 | $0.50 |
418
+ | `kilo/z-ai/glm-5.3-flashx` | 1.0M | | | | | | $0.37 | $1 |
419
419
  | `kilo/z-ai/glm-5v-turbo` | 203K | | | | | | $1 | $4 |
420
420
 
421
421
  Model availability, capabilities, context windows, and pricing are sourced from [models.dev](https://models.dev) and may change.
@@ -4,7 +4,7 @@
4
4
 
5
5
  # ![LLM Gateway logo](https://models.dev/logos/llmgateway-providers.svg)LLM Gateway
6
6
 
7
- Access 404 LLM Gateway models through Mastra's model router. Authentication is handled automatically using the `LLMGATEWAY_API_KEY` environment variable.
7
+ Access 402 LLM Gateway models through Mastra's model router. Authentication is handled automatically using the `LLMGATEWAY_API_KEY` environment variable.
8
8
 
9
9
  Learn more in the [LLM Gateway documentation](https://llmgateway.io/docs).
10
10
 
@@ -415,8 +415,6 @@ for await (const chunk of stream) {
415
415
  | `llmgateway-providers/vertex-openai/qwen3-coder-480b-a35b-instruct` | 262K | | | | | | $0.22 | $2 |
416
416
  | `llmgateway-providers/vertex-openai/qwen3-next-80b-a3b-instruct` | 131K | | | | | | $0.15 | $1 |
417
417
  | `llmgateway-providers/vertex-openai/qwen3-next-80b-a3b-thinking` | 131K | | | | | | $0.15 | $1 |
418
- | `llmgateway-providers/vichar-ai/glm-5.3` | 1.0M | | | | | | $1 | $4 |
419
- | `llmgateway-providers/vichar-ai/glm-5.3-flash` | 1.0M | | | | | | $0.15 | $0.50 |
420
418
  | `llmgateway-providers/xai/grok-4` | 256K | | | | | | $3 | $15 |
421
419
  | `llmgateway-providers/xai/grok-4-20-beta-0309-non-reasoning` | 2.0M | | | | | | $2 | $6 |
422
420
  | `llmgateway-providers/xai/grok-4-20-beta-0309-reasoning` | 2.0M | | | | | | $2 | $6 |
@@ -2,11 +2,11 @@
2
2
 
3
3
  > Discover all available pages from the documentation index: https://mastra.ai/llms.txt
4
4
 
5
- # ![MiniMax Token Plan (minimaxi.com) logo](https://models.dev/logos/minimax-cn-coding-plan.svg)MiniMax Token Plan (minimaxi.com)
5
+ # ![MiniMax Token Plan (minimax.cn) logo](https://models.dev/logos/minimax-cn-coding-plan.svg)MiniMax Token Plan (minimax.cn)
6
6
 
7
- Access 7 MiniMax Token Plan (minimaxi.com) models through Mastra's model router. Authentication is handled automatically using the `MINIMAX_API_KEY` environment variable.
7
+ Access 7 MiniMax Token Plan (minimax.cn) models through Mastra's model router. Authentication is handled automatically using the `MINIMAX_API_KEY` environment variable.
8
8
 
9
- Learn more in the [MiniMax Token Plan (minimaxi.com) documentation](https://platform.minimaxi.com/docs/token-plan/intro).
9
+ Learn more in the [MiniMax Token Plan (minimax.cn) documentation](https://platform.minimaxi.com/docs/token-plan/intro).
10
10
 
11
11
  ```bash
12
12
  MINIMAX_API_KEY=your-api-key
@@ -32,7 +32,7 @@ for await (const chunk of stream) {
32
32
  }
33
33
  ```
34
34
 
35
- > **Note:** Mastra uses the OpenAI-compatible `/chat/completions` endpoint. Some provider-specific features may not be available. Check the [MiniMax Token Plan (minimaxi.com) documentation](https://platform.minimaxi.com/docs/token-plan/intro) for details.
35
+ > **Note:** Mastra uses the OpenAI-compatible `/chat/completions` endpoint. Some provider-specific features may not be available. Check the [MiniMax Token Plan (minimax.cn) documentation](https://platform.minimaxi.com/docs/token-plan/intro) for details.
36
36
 
37
37
  ## Models
38
38
 
@@ -57,7 +57,7 @@ const agent = new Agent({
57
57
  id: "custom-agent",
58
58
  name: "custom-agent",
59
59
  model: {
60
- url: "https://api.minimaxi.com/anthropic/v1",
60
+ url: "https://api.minimax.cn/anthropic/v1",
61
61
  id: "minimax-cn-coding-plan/MiniMax-M2",
62
62
  apiKey: process.env.MINIMAX_API_KEY,
63
63
  headers: {
@@ -2,11 +2,11 @@
2
2
 
3
3
  > Discover all available pages from the documentation index: https://mastra.ai/llms.txt
4
4
 
5
- # ![MiniMax (minimaxi.com) logo](https://models.dev/logos/minimax-cn.svg)MiniMax (minimaxi.com)
5
+ # ![MiniMax (minimax.cn) logo](https://models.dev/logos/minimax-cn.svg)MiniMax (minimax.cn)
6
6
 
7
- Access 7 MiniMax (minimaxi.com) models through Mastra's model router. Authentication is handled automatically using the `MINIMAX_API_KEY` environment variable.
7
+ Access 7 MiniMax (minimax.cn) models through Mastra's model router. Authentication is handled automatically using the `MINIMAX_API_KEY` environment variable.
8
8
 
9
- Learn more in the [MiniMax (minimaxi.com) documentation](https://platform.minimaxi.com/docs/guides/quickstart).
9
+ Learn more in the [MiniMax (minimax.cn) documentation](https://platform.minimaxi.com/docs/guides/quickstart).
10
10
 
11
11
  ```bash
12
12
  MINIMAX_API_KEY=your-api-key
@@ -32,7 +32,7 @@ for await (const chunk of stream) {
32
32
  }
33
33
  ```
34
34
 
35
- > **Note:** Mastra uses the OpenAI-compatible `/chat/completions` endpoint. Some provider-specific features may not be available. Check the [MiniMax (minimaxi.com) documentation](https://platform.minimaxi.com/docs/guides/quickstart) for details.
35
+ > **Note:** Mastra uses the OpenAI-compatible `/chat/completions` endpoint. Some provider-specific features may not be available. Check the [MiniMax (minimax.cn) documentation](https://platform.minimaxi.com/docs/guides/quickstart) for details.
36
36
 
37
37
  ## Models
38
38
 
@@ -57,7 +57,7 @@ const agent = new Agent({
57
57
  id: "custom-agent",
58
58
  name: "custom-agent",
59
59
  model: {
60
- url: "https://api.minimaxi.com/anthropic/v1",
60
+ url: "https://api.minimax.cn/anthropic/v1",
61
61
  id: "minimax-cn/MiniMax-M2",
62
62
  apiKey: process.env.MINIMAX_API_KEY,
63
63
  headers: {
@@ -4,7 +4,7 @@
4
4
 
5
5
  # ![NanoGPT logo](https://models.dev/logos/nano-gpt.svg)NanoGPT
6
6
 
7
- Access 572 NanoGPT models through Mastra's model router. Authentication is handled automatically using the `NANO_GPT_API_KEY` environment variable.
7
+ Access 577 NanoGPT models through Mastra's model router. Authentication is handled automatically using the `NANO_GPT_API_KEY` environment variable.
8
8
 
9
9
  Learn more in the [NanoGPT documentation](https://docs.nano-gpt.com).
10
10
 
@@ -221,6 +221,7 @@ for await (const chunk of stream) {
221
221
  | `nano-gpt/glm-4.1v-thinking-flashx` | 64K | | | | | | $0.30 | $0.30 |
222
222
  | `nano-gpt/GLM-4.6-Derestricted-v5` | 131K | | | | | | $0.40 | $2 |
223
223
  | `nano-gpt/glm-z1-airx` | 32K | | | | | | $0.70 | $0.70 |
224
+ | `nano-gpt/google/diffusiongemma` | 262K | | | | | | $0.05 | $0.15 |
224
225
  | `nano-gpt/google/gemini-3-flash-preview` | 1.0M | | | | | | $0.50 | $3 |
225
226
  | `nano-gpt/google/gemini-3-flash-preview-thinking` | 1.0M | | | | | | $0.50 | $3 |
226
227
  | `nano-gpt/google/gemini-3.1-flash-lite` | 1.0M | | | | | | $0.25 | $2 |
@@ -238,9 +239,11 @@ for await (const chunk of stream) {
238
239
  | `nano-gpt/google/gemini-flash-lite-latest` | 1.0M | | | | | | $0.30 | $3 |
239
240
  | `nano-gpt/google/gemini-pro-latest` | 1.0M | | | | | | $2 | $12 |
240
241
  | `nano-gpt/google/gemma-4-26b-a4b-it` | 262K | | | | | | $0.12 | $0.38 |
242
+ | `nano-gpt/google/gemma-4-26b-a4b-it-cybersecurity` | 262K | | | | | | $0.11 | $0.33 |
241
243
  | `nano-gpt/google/gemma-4-26b-a4b-it:thinking` | 262K | | | | | | $0.13 | $0.40 |
242
244
  | `nano-gpt/google/gemma-4-31b-it` | 262K | | | | | | $0.10 | $0.45 |
243
245
  | `nano-gpt/google/gemma-4-31b-it:thinking` | 262K | | | | | | $0.10 | $0.35 |
246
+ | `nano-gpt/google/gemma4-31b-splituntied` | 262K | | | | | | $0.10 | $0.30 |
244
247
  | `nano-gpt/Gryphe/MythoMax-L2-13b` | 4K | | | | | | $0.10 | $0.10 |
245
248
  | `nano-gpt/hermes-high` | 1.0M | | | | | | $1 | $3 |
246
249
  | `nano-gpt/hermes-low` | 1.0M | | | | | | $1 | $3 |
@@ -350,6 +353,7 @@ for await (const chunk of stream) {
350
353
  | `nano-gpt/nvidia/nemotron-3-super-120b-a12b:thinking` | 262K | | | | | | $0.05 | $0.25 |
351
354
  | `nano-gpt/nvidia/nemotron-3-ultra-550b-a55b` | 1.0M | | | | | | $0.50 | $3 |
352
355
  | `nano-gpt/nvidia/nemotron-3-ultra-550b-a55b:thinking` | 1.0M | | | | | | $0.50 | $3 |
356
+ | `nano-gpt/nvidia/nemotron-3.5-content-safety` | 131K | | | | | | $0.05 | $0.15 |
353
357
  | `nano-gpt/nvidia/nemotron-3.5-lightning` | 1.0M | | | | | | $0.05 | $0.20 |
354
358
  | `nano-gpt/nvidia/nemotron-3.5-lightning:thinking` | 1.0M | | | | | | $0.05 | $0.20 |
355
359
  | `nano-gpt/openai/gpt-3.5-turbo` | 16K | | | | | | $0.50 | $2 |
@@ -467,12 +471,13 @@ for await (const chunk of stream) {
467
471
  | `nano-gpt/qwen/qwen3.7-plus` | 992K | | | | | | $0.40 | $2 |
468
472
  | `nano-gpt/qwen/qwen3.7-plus:thinking` | 984K | | | | | | $0.40 | $2 |
469
473
  | `nano-gpt/qwen/qwen3.8-27b` | 262K | | | | | | $0.15 | $0.70 |
474
+ | `nano-gpt/qwen/qwen3.8-27b-cybersecurity` | 262K | | | | | | $0.10 | $0.60 |
470
475
  | `nano-gpt/qwen/qwen3.8-27b-fable` | 524K | | | | | | $0.25 | $2 |
471
476
  | `nano-gpt/qwen/qwen3.8-27b-obliterated` | 524K | | | | | | $0.25 | $2 |
472
477
  | `nano-gpt/qwen/qwen3.8-27b-obliterated:thinking` | 524K | | | | | | $0.25 | $2 |
473
478
  | `nano-gpt/qwen/qwen3.8-27b-queen` | 524K | | | | | | $0.25 | $2 |
474
- | `nano-gpt/qwen/qwen3.8-27b-uncensored` | 524K | | | | | | $0.25 | $2 |
475
- | `nano-gpt/qwen/qwen3.8-27b-uncensored:thinking` | 524K | | | | | | $0.25 | $2 |
479
+ | `nano-gpt/qwen/qwen3.8-27b-uncensored` | 524K | | | | | | $0.20 | $2 |
480
+ | `nano-gpt/qwen/qwen3.8-27b-uncensored:thinking` | 524K | | | | | | $0.20 | $2 |
476
481
  | `nano-gpt/qwen/qwen3.8-27b:thinking` | 262K | | | | | | $0.15 | $0.70 |
477
482
  | `nano-gpt/qwen/qwen3.8-flash` | 992K | | | | | | $0.14 | $0.42 |
478
483
  | `nano-gpt/qwen/qwen3.8-max` | 991K | | | | | | $2 | $6 |
@@ -494,7 +499,6 @@ for await (const chunk of stream) {
494
499
  | `nano-gpt/sarvam-105b` | 131K | | | | | | $0.05 | $0.21 |
495
500
  | `nano-gpt/shisa-ai/shisa-v2-llama3.3-70b` | 128K | | | | | | $0.50 | $0.50 |
496
501
  | `nano-gpt/shisa-ai/shisa-v2.1-llama3.3-70b` | 33K | | | | | | $0.50 | $0.50 |
497
- | `nano-gpt/slowburn/gemma4-31b-splituntied` | 262K | | | | | | $0.10 | $0.30 |
498
502
  | `nano-gpt/soob3123/amoral-gemma3-27B-v2` | 33K | | | | | | $0.30 | $0.30 |
499
503
  | `nano-gpt/soob3123/GrayLine-Qwen3-8B` | 33K | | | | | | $0.30 | $0.30 |
500
504
  | `nano-gpt/soob3123/Veiled-Calla-12B` | 33K | | | | | | $0.30 | $0.30 |
@@ -605,6 +609,7 @@ for await (const chunk of stream) {
605
609
  | `nano-gpt/z-ai/glm-5.2:thinking` | 1.0M | | | | | | $0.42 | $1 |
606
610
  | `nano-gpt/z-ai/glm-5.3` | 1.0M | | | | | | $1 | $3 |
607
611
  | `nano-gpt/z-ai/glm-5.3-flash` | 1.0M | | | | | | $0.07 | $0.25 |
612
+ | `nano-gpt/z-ai/glm-5.3-flash-cybersecurity` | 1.0M | | | | | | $0.15 | $0.50 |
608
613
  | `nano-gpt/z-ai/glm-5.3-flash-uncensored` | 1.0M | | | | | | $0.20 | $0.80 |
609
614
  | `nano-gpt/z-ai/glm-5.3:thinking` | 1.0M | | | | | | $1 | $3 |
610
615
  | `nano-gpt/z-ai/glm-5v-turbo` | 203K | | | | | | $1 | $4 |
@@ -4,7 +4,7 @@
4
4
 
5
5
  # ![Nebius Token Factory logo](https://models.dev/logos/nebius.svg)Nebius Token Factory
6
6
 
7
- Access 17 Nebius Token Factory models through Mastra's model router. Authentication is handled automatically using the `NEBIUS_API_KEY` environment variable.
7
+ Access 20 Nebius Token Factory models through Mastra's model router. Authentication is handled automatically using the `NEBIUS_API_KEY` environment variable.
8
8
 
9
9
  Learn more in the [Nebius Token Factory documentation](https://docs.tokenfactory.nebius.com/).
10
10
 
@@ -40,6 +40,8 @@ for await (const chunk of stream) {
40
40
  | ------------------------------------------- | ------- | ----- | --------- | ----- | ----- | ----- | ---------- | ----------- |
41
41
  | `nebius/deepseek-ai/DeepSeek-V4-Flash-0731` | 1.0M | | | | | | $0.14 | $0.28 |
42
42
  | `nebius/deepseek-ai/DeepSeek-V4-Pro` | 1.0M | | | | | | $2 | $4 |
43
+ | `nebius/deepseek-ai/DeepSeek-V4-Pro-0813` | 979K | | | | | | $1 | $4 |
44
+ | `nebius/deepseek-ai/DeepSeek-V4.1-Flash` | 1.0M | | | | | | $0.30 | $1 |
43
45
  | `nebius/google/gemma-3-27b-it` | 110K | | | | | | $0.10 | $0.30 |
44
46
  | `nebius/MiniMaxAI/MiniMax-M3` | 1.0M | | | | | | $0.30 | $1 |
45
47
  | `nebius/moonshotai/Kimi-K2.7-Code` | 262K | | | | | | $0.95 | $4 |
@@ -54,6 +56,7 @@ for await (const chunk of stream) {
54
56
  | `nebius/Qwen/Qwen3-Embedding-8B` | 41K | | | | | | $0.01 | — |
55
57
  | `nebius/Qwen/Qwen3.5-397B-A17B` | 262K | | | | | | $0.60 | $4 |
56
58
  | `nebius/zai-org/GLM-5.2` | 1.0M | | | | | | $1 | $4 |
59
+ | `nebius/zai-org/GLM-5.3` | 1.0M | | | | | | $1 | $4 |
57
60
  | `nebius/zai-org/GLM-5.3-Flash` | 1.0M | | | | | | $0.15 | $0.50 |
58
61
 
59
62
  Model availability, capabilities, context windows, and pricing are sourced from [models.dev](https://models.dev) and may change.
@@ -4,7 +4,7 @@
4
4
 
5
5
  # ![OpenCode Zen logo](https://models.dev/logos/opencode.svg)OpenCode Zen
6
6
 
7
- Access 103 OpenCode Zen models through Mastra's model router. Authentication is handled automatically using the `OPENCODE_API_KEY` environment variable.
7
+ Access 107 OpenCode Zen models through Mastra's model router. Authentication is handled automatically using the `OPENCODE_API_KEY` environment variable.
8
8
 
9
9
  Learn more in the [OpenCode Zen documentation](https://opencode.ai/docs/zen).
10
10
 
@@ -54,6 +54,7 @@ for await (const chunk of stream) {
54
54
  | `opencode/deepseek-v4-flash` | 1.0M | | | | | | $0.14 | $0.28 |
55
55
  | `opencode/deepseek-v4-flash-vision-exp` | 1.0M | | | | | | $0.14 | $0.28 |
56
56
  | `opencode/deepseek-v4-pro` | 1.0M | | | | | | $2 | $4 |
57
+ | `opencode/deepseek-v4.1-flash` | 1.0M | | | | | | $0.30 | $1 |
57
58
  | `opencode/gemini-3-flash` | 1.0M | | | | | | $0.50 | $3 |
58
59
  | `opencode/gemini-3.1-pro` | 1.0M | | | | | | $2 | $12 |
59
60
  | `opencode/gemini-3.5-flash` | 1.0M | | | | | | $2 | $9 |
@@ -90,6 +91,8 @@ for await (const chunk of stream) {
90
91
  | `opencode/grok-4.5` | 500K | | | | | | $2 | $6 |
91
92
  | `opencode/grok-4.6` | 500K | | | | | | $2 | $6 |
92
93
  | `opencode/grok-build-0.1` | 256K | | | | | | $1 | $2 |
94
+ | `opencode/jev-1.13` | 64K | | | | | | $0.04 | — |
95
+ | `opencode/jev-1.13-free` | 64K | | | | | | — | — |
93
96
  | `opencode/jev-latest` | 64K | | | | | | $0.04 | — |
94
97
  | `opencode/kimi-k2.5` | 262K | | | | | | $0.60 | $3 |
95
98
  | `opencode/kimi-k2.6` | 262K | | | | | | $0.95 | $4 |
@@ -108,6 +111,7 @@ for await (const chunk of stream) {
108
111
  | `opencode/nemotron-3.5-lightning-free` | 262K | | | | | | — | — |
109
112
  | `opencode/qwen3.5-plus` | 262K | | | | | | $0.20 | $1 |
110
113
  | `opencode/qwen3.6-plus` | 262K | | | | | | $0.50 | $3 |
114
+ | `opencode/qwen3.8-flash` | 1.0M | | | | | | $0.15 | $0.47 |
111
115
 
112
116
  Model availability, capabilities, context windows, and pricing are sourced from [models.dev](https://models.dev) and may change.
113
117
 
@@ -4,7 +4,7 @@
4
4
 
5
5
  # ![TensorX logo](https://models.dev/logos/tensorx.svg)TensorX
6
6
 
7
- Access 26 TensorX models through Mastra's model router. Authentication is handled automatically using the `TENSORX_API_KEY` environment variable.
7
+ Access 25 TensorX models through Mastra's model router. Authentication is handled automatically using the `TENSORX_API_KEY` environment variable.
8
8
 
9
9
  Learn more in the [TensorX documentation](https://docs.tensorx.ai/).
10
10
 
@@ -19,7 +19,7 @@ const agent = new Agent({
19
19
  id: "my-agent",
20
20
  name: "My Agent",
21
21
  instructions: "You are a helpful assistant",
22
- model: "tensorx/deepseek/deepseek-chat-v3.1"
22
+ model: "tensorx/deepseek/deepseek-r1-0528"
23
23
  });
24
24
 
25
25
  // Generate a response
@@ -36,34 +36,33 @@ for await (const chunk of stream) {
36
36
 
37
37
  ## Models
38
38
 
39
- | Model | Context | Tools | Reasoning | Image | Audio | Video | Input $/1M | Output $/1M |
40
- | ------------------------------------------- | ------- | ----- | --------- | ----- | ----- | ----- | ---------- | ----------- |
41
- | `tensorx/deepseek/deepseek-chat-v3.1` | 164K | | | | | | $0.20 | $0.80 |
42
- | `tensorx/deepseek/deepseek-r1-0528` | 164K | | | | | | $0.66 | $3 |
43
- | `tensorx/deepseek/deepseek-v3.2` | 164K | | | | | | $0.30 | $0.50 |
44
- | `tensorx/deepseek/deepseek-v4-flash` | 1.0M | | | | | | $0.15 | $0.30 |
45
- | `tensorx/deepseek/deepseek-v4-flash-0731` | 1.0M | | | | | | $0.25 | $0.30 |
46
- | `tensorx/deepseek/deepseek-v4-pro` | 1.0M | | | | | | $2 | $4 |
47
- | `tensorx/deepseek/deepseek-v4.1-flash` | 1.0M | | | | | | $0.50 | $2 |
48
- | `tensorx/minimax/minimax-m2.5` | 197K | | | | | | $0.30 | $1 |
49
- | `tensorx/minimax/minimax-m3` | 1.0M | | | | | | $0.40 | $2 |
50
- | `tensorx/moonshotai/kimi-k2.5` | 262K | | | | | | $0.50 | $3 |
51
- | `tensorx/moonshotai/kimi-k2.6` | 262K | | | | | | $1 | $4 |
52
- | `tensorx/moonshotai/kimi-k2.7-code` | 262K | | | | | | $1 | $5 |
53
- | `tensorx/moonshotai/kimi-k3` | 1.0M | | | | | | $3 | $15 |
54
- | `tensorx/nvidia/nemotron-3-super-120b-a12b` | 262K | | | | | | $0.30 | $0.90 |
55
- | `tensorx/openai/gpt-oss-120b` | 131K | | | | | | $0.04 | $0.20 |
56
- | `tensorx/qwen/qwen3-235b-a22b-2507` | 131K | | | | | | $0.07 | $0.46 |
57
- | `tensorx/qwen/qwen3-coder-30b-a3b-instruct` | 262K | | | | | | $0.06 | $0.25 |
58
- | `tensorx/qwen/qwen3-vl-235b-a22b-instruct` | 131K | | | | | | $0.21 | $2 |
59
- | `tensorx/qwen/qwen3.5-122b-a10b` | 262K | | | | | | $0.50 | $4 |
60
- | `tensorx/qwen/qwen3.5-9b` | 262K | | | | | | $0.15 | $0.20 |
61
- | `tensorx/z-ai/glm-4.7` | 200K | | | | | | $0.60 | $2 |
62
- | `tensorx/z-ai/glm-5` | 203K | | | | | | $1 | $3 |
63
- | `tensorx/z-ai/glm-5-turbo` | 203K | | | | | | $1 | $4 |
64
- | `tensorx/z-ai/glm-5.1` | 203K | | | | | | $1 | $4 |
65
- | `tensorx/z-ai/glm-5.2` | 1.0M | | | | | | $2 | $5 |
66
- | `tensorx/z-ai/glm-5v-turbo` | 203K | | | | | | $1 | $4 |
39
+ | Model | Context | Tools | Reasoning | Image | Audio | Video | Input $/1M | Output $/1M |
40
+ | ----------------------------------------- | ------- | ----- | --------- | ----- | ----- | ----- | ---------- | ----------- |
41
+ | `tensorx/deepseek/deepseek-r1-0528` | 164K | | | | | | $0.66 | $3 |
42
+ | `tensorx/deepseek/deepseek-v3.2` | 164K | | | | | | $0.30 | $0.50 |
43
+ | `tensorx/deepseek/deepseek-v4-flash-0731` | 1.0M | | | | | | $0.25 | $0.30 |
44
+ | `tensorx/deepseek/deepseek-v4-pro` | 1.0M | | | | | | $2 | $4 |
45
+ | `tensorx/deepseek/deepseek-v4-pro-0813` | 1.0M | | | | | | $2 | $4 |
46
+ | `tensorx/deepseek/deepseek-v4.1-flash` | 1.0M | | | | | | $0.50 | $2 |
47
+ | `tensorx/minimax/minimax-m2.5` | 197K | | | | | | $0.30 | $1 |
48
+ | `tensorx/minimax/minimax-m3` | 1.0M | | | | | | $0.40 | $2 |
49
+ | `tensorx/moonshotai/kimi-k2.5` | 262K | | | | | | $0.50 | $3 |
50
+ | `tensorx/moonshotai/kimi-k2.6` | 262K | | | | | | $1 | $4 |
51
+ | `tensorx/moonshotai/kimi-k2.7-code` | 262K | | | | | | $1 | $5 |
52
+ | `tensorx/moonshotai/kimi-k3` | 1.0M | | | | | | $3 | $15 |
53
+ | `tensorx/qwen/qwen3-235b-a22b-2507` | 131K | | | | | | $0.07 | $0.46 |
54
+ | `tensorx/qwen/qwen3.5-122b-a10b` | 262K | | | | | | $0.50 | $4 |
55
+ | `tensorx/qwen/qwen3.5-9b` | 262K | | | | | | $0.15 | $0.20 |
56
+ | `tensorx/qwen/qwen3.8-2.4t-a95b` | 262K | | | | | | $3 | $6 |
57
+ | `tensorx/qwen/qwen3.8-27b` | 262K | | | | | | $0.40 | $2 |
58
+ | `tensorx/qwen/qwen3.8-flash-next` | 262K | | | | | | $0.20 | $0.50 |
59
+ | `tensorx/z-ai/glm-5` | 203K | | | | | | $1 | $3 |
60
+ | `tensorx/z-ai/glm-5-turbo` | 203K | | | | | | $1 | $4 |
61
+ | `tensorx/z-ai/glm-5.1` | 203K | | | | | | $1 | $4 |
62
+ | `tensorx/z-ai/glm-5.2` | 1.0M | | | | | | $2 | $5 |
63
+ | `tensorx/z-ai/glm-5.3` | 1.0M | | | | | | $2 | $5 |
64
+ | `tensorx/z-ai/glm-5.3-flash` | 1.0M | | | | | | $0.20 | $0.50 |
65
+ | `tensorx/z-ai/glm-5v-turbo` | 203K | | | | | | $1 | $4 |
67
66
 
68
67
  Model availability, capabilities, context windows, and pricing are sourced from [models.dev](https://models.dev) and may change.
69
68
 
@@ -77,7 +76,7 @@ const agent = new Agent({
77
76
  name: "custom-agent",
78
77
  model: {
79
78
  url: "https://api.tensorx.ai/v1",
80
- id: "tensorx/deepseek/deepseek-chat-v3.1",
79
+ id: "tensorx/deepseek/deepseek-r1-0528",
81
80
  apiKey: process.env.TENSORX_API_KEY,
82
81
  headers: {
83
82
  "X-Custom-Header": "value"
@@ -96,7 +95,7 @@ const agent = new Agent({
96
95
  const useAdvanced = requestContext.task === "complex";
97
96
  return useAdvanced
98
97
  ? "tensorx/z-ai/glm-5v-turbo"
99
- : "tensorx/deepseek/deepseek-chat-v3.1";
98
+ : "tensorx/deepseek/deepseek-r1-0528";
100
99
  }
101
100
  });
102
101
  ```
@@ -4,7 +4,7 @@
4
4
 
5
5
  # ![Z.AI logo](https://models.dev/logos/zai.svg)Z.AI
6
6
 
7
- Access 16 Z.AI models through Mastra's model router. Authentication is handled automatically using the `ZHIPU_API_KEY` environment variable.
7
+ Access 17 Z.AI models through Mastra's model router. Authentication is handled automatically using the `ZHIPU_API_KEY` environment variable.
8
8
 
9
9
  Learn more in the [Z.AI documentation](https://docs.z.ai/guides/overview/pricing).
10
10
 
@@ -52,7 +52,8 @@ for await (const chunk of stream) {
52
52
  | `zai/glm-5.1` | 200K | | | | | | $1 | $4 |
53
53
  | `zai/glm-5.2` | 1.0M | | | | | | $1 | $4 |
54
54
  | `zai/glm-5.3` | 1.0M | | | | | | $1 | $4 |
55
- | `zai/glm-5.3-flash` | 1.0M | | | | | | $0.07 | $0.25 |
55
+ | `zai/glm-5.3-flash` | 1.0M | | | | | | $0.15 | $0.50 |
56
+ | `zai/glm-5.3-flashx` | 1.0M | | | | | | $0.37 | $1 |
56
57
  | `zai/glm-5v-turbo` | 200K | | | | | | $1 | $4 |
57
58
 
58
59
  Model availability, capabilities, context windows, and pricing are sourced from [models.dev](https://models.dev) and may change.
@@ -4,7 +4,7 @@
4
4
 
5
5
  # ![ZenMux logo](https://models.dev/logos/zenmux.svg)ZenMux
6
6
 
7
- Access 120 ZenMux models through Mastra's model router. Authentication is handled automatically using the `ZENMUX_API_KEY` environment variable.
7
+ Access 122 ZenMux models through Mastra's model router. Authentication is handled automatically using the `ZENMUX_API_KEY` environment variable.
8
8
 
9
9
  Learn more in the [ZenMux documentation](https://docs.zenmux.ai).
10
10
 
@@ -117,31 +117,38 @@ for await (const chunk of stream) {
117
117
  | `zenmux/volcengine/doubao-seed-2.0-lite` | 256K | | | | | | $0.09 | $0.51 |
118
118
  | `zenmux/volcengine/doubao-seed-2.0-mini` | 256K | | | | | | $0.03 | $0.28 |
119
119
  | `zenmux/volcengine/doubao-seed-2.0-pro` | 256K | | | | | | $0.45 | $2 |
120
- | `zenmux/x-ai/grok-4.2-fast` | 2.0M | | | | | | $3 | $9 |
121
- | `zenmux/x-ai/grok-4.2-fast-non-reasoning` | 2.0M | | | | | | $3 | $9 |
120
+ | `zenmux/x-ai/grok-4.2-fast` | 2.0M | | | | | | $2 | $6 |
121
+ | `zenmux/x-ai/grok-4.2-fast-non-reasoning` | 2.0M | | | | | | $2 | $6 |
122
122
  | `zenmux/x-ai/grok-4.3` | 1.0M | | | | | | $1 | $3 |
123
123
  | `zenmux/x-ai/grok-4.5` | 500K | | | | | | $2 | $6 |
124
+ | `zenmux/x-ai/grok-4.6` | 500K | | | | | | $2 | $6 |
124
125
  | `zenmux/x-ai/grok-build-0.1` | 256K | | | | | | $1 | $2 |
126
+ | `zenmux/x-ai/grok-imagine-image-2.0` | 66K | | | | | | — | — |
127
+ | `zenmux/x-ai/grok-voice-stt-1.0` | 15K | | | | | | — | — |
128
+ | `zenmux/x-ai/grok-voice-tts-1.0` | 15K | | | | | | — | — |
125
129
  | `zenmux/xiaomi/mimo-v2-flash` | 262K | | | | | | $0.10 | $0.30 |
126
130
  | `zenmux/xiaomi/mimo-v2-omni` | 265K | | | | | | $0.40 | $2 |
127
131
  | `zenmux/xiaomi/mimo-v2-pro` | 1.0M | | | | | | $1 | $3 |
128
132
  | `zenmux/xiaomi/mimo-v2.5` | 1.0M | | | | | | $0.40 | $2 |
129
133
  | `zenmux/xiaomi/mimo-v2.5-pro` | 1.0M | | | | | | $1 | $3 |
130
- | `zenmux/z-ai/glm-4.5` | 128K | | | | | | $0.35 | $2 |
131
- | `zenmux/z-ai/glm-4.5-air` | 128K | | | | | | $0.11 | $0.56 |
132
- | `zenmux/z-ai/glm-4.6` | 200K | | | | | | $0.35 | $2 |
133
- | `zenmux/z-ai/glm-4.6v` | 200K | | | | | | $0.14 | $0.42 |
134
- | `zenmux/z-ai/glm-4.6v-flash` | 200K | | | | | | $0.02 | $0.21 |
134
+ | `zenmux/z-ai/glm-4.5` | 128K | | | | | | $0.29 | $1 |
135
+ | `zenmux/z-ai/glm-4.5-air` | 128K | | | | | | $0.12 | $0.29 |
136
+ | `zenmux/z-ai/glm-4.6` | 200K | | | | | | $0.29 | $1 |
137
+ | `zenmux/z-ai/glm-4.6v` | 200K | | | | | | $0.15 | $0.44 |
138
+ | `zenmux/z-ai/glm-4.6v-flash` | 200K | | | | | | $0.02 | $0.22 |
135
139
  | `zenmux/z-ai/glm-4.6v-flash-free` | 200K | | | | | | — | — |
136
- | `zenmux/z-ai/glm-4.7` | 200K | | | | | | $0.28 | $1 |
140
+ | `zenmux/z-ai/glm-4.7` | 200K | | | | | | $0.29 | $1 |
137
141
  | `zenmux/z-ai/glm-4.7-flash-free` | 200K | | | | | | — | — |
138
- | `zenmux/z-ai/glm-4.7-flashx` | 200K | | | | | | $0.07 | $0.42 |
142
+ | `zenmux/z-ai/glm-4.7-flashx` | 200K | | | | | | $0.07 | $0.44 |
139
143
  | `zenmux/z-ai/glm-5` | 200K | | | | | | $0.58 | $3 |
140
- | `zenmux/z-ai/glm-5-turbo` | 200K | | | | | | $0.88 | $3 |
144
+ | `zenmux/z-ai/glm-5-turbo` | 200K | | | | | | $0.73 | $3 |
141
145
  | `zenmux/z-ai/glm-5.1` | 200K | | | | | | $0.88 | $4 |
142
- | `zenmux/z-ai/glm-5.2` | 1.0M | | | | | | $1 | $5 |
143
- | `zenmux/z-ai/glm-5.2-free` | 1.0M | | | | | | — | — |
146
+ | `zenmux/z-ai/glm-5.2` | 1.0M | | | | | | $0.98 | $3 |
147
+ | `zenmux/z-ai/glm-5.3` | 1.0M | | | | | | $1 | $4 |
148
+ | `zenmux/z-ai/glm-5.3-flash` | 1.0M | | | | | | $0.15 | $0.50 |
149
+ | `zenmux/z-ai/glm-5.3-flashx` | 1.0M | | | | | | $0.38 | $1 |
144
150
  | `zenmux/z-ai/glm-5v-turbo` | 200K | | | | | | $0.73 | $3 |
151
+ | `zenmux/z-ai/glm-image` | 10K | | | | | | — | — |
145
152
 
146
153
  Model availability, capabilities, context windows, and pricing are sourced from [models.dev](https://models.dev) and may change.
147
154
 
@@ -173,7 +180,7 @@ const agent = new Agent({
173
180
  model: ({ requestContext }) => {
174
181
  const useAdvanced = requestContext.task === "complex";
175
182
  return useAdvanced
176
- ? "zenmux/z-ai/glm-5v-turbo"
183
+ ? "zenmux/z-ai/glm-image"
177
184
  : "zenmux/anthropic/claude-3.5-haiku";
178
185
  }
179
186
  });
@@ -4,7 +4,7 @@
4
4
 
5
5
  # ![Zhipu AI logo](https://models.dev/logos/zhipuai.svg)Zhipu AI
6
6
 
7
- Access 15 Zhipu AI models through Mastra's model router. Authentication is handled automatically using the `ZHIPU_API_KEY` environment variable.
7
+ Access 16 Zhipu AI models through Mastra's model router. Authentication is handled automatically using the `ZHIPU_API_KEY` environment variable.
8
8
 
9
9
  Learn more in the [Zhipu AI documentation](https://docs.z.ai/guides/overview/pricing).
10
10
 
@@ -51,7 +51,8 @@ for await (const chunk of stream) {
51
51
  | `zhipuai/glm-5.1` | 200K | | | | | | $1 | $4 |
52
52
  | `zhipuai/glm-5.2` | 1.0M | | | | | | $1 | $4 |
53
53
  | `zhipuai/glm-5.3` | 1.0M | | | | | | $1 | $4 |
54
- | `zhipuai/glm-5.3-flash` | 1.0M | | | | | | $0.07 | $0.25 |
54
+ | `zhipuai/glm-5.3-flash` | 1.0M | | | | | | $0.15 | $0.50 |
55
+ | `zhipuai/glm-5.3-flashx` | 1.0M | | | | | | $0.37 | $1 |
55
56
  | `zhipuai/glm-5v-turbo` | 200K | | | | | | $5 | $22 |
56
57
 
57
58
  Model availability, capabilities, context windows, and pricing are sourced from [models.dev](https://models.dev) and may change.
@@ -111,10 +111,10 @@ Direct access to individual AI model providers. Each provider offers unique mode
111
111
  - [Meganova](https://mastra.ai/models/providers/meganova)
112
112
  - [Melious](https://mastra.ai/models/providers/melious)
113
113
  - [Meta](https://mastra.ai/models/providers/meta)
114
+ - [MiniMax (minimax.cn)](https://mastra.ai/models/providers/minimax-cn)
114
115
  - [MiniMax (minimax.io)](https://mastra.ai/models/providers/minimax)
115
- - [MiniMax (minimaxi.com)](https://mastra.ai/models/providers/minimax-cn)
116
+ - [MiniMax Token Plan (minimax.cn)](https://mastra.ai/models/providers/minimax-cn-coding-plan)
116
117
  - [MiniMax Token Plan (minimax.io)](https://mastra.ai/models/providers/minimax-coding-plan)
117
- - [MiniMax Token Plan (minimaxi.com)](https://mastra.ai/models/providers/minimax-cn-coding-plan)
118
118
  - [Mixlayer](https://mastra.ai/models/providers/mixlayer)
119
119
  - [Moark](https://mastra.ai/models/providers/moark)
120
120
  - [Modal](https://mastra.ai/models/providers/modal)
@@ -168,7 +168,7 @@ const result = await agent.generate('message for agent')
168
168
 
169
169
  **options.modelSettings.frequencyPenalty** (`number`): Penalty for token frequency (-2 to 2). Reduces repetition of frequent tokens.
170
170
 
171
- **options.modelSettings.timeout** (`object`): Time-based execution budget for the run. Accepts totalMs, the maximum duration of the entire agent run across every loop iteration, tool call and retry, and stepMs, the maximum duration of a single model call including the time spent consuming its stream. Exceeding either budget fails with a MastraTimeoutError. A totalMs timeout ends the run and does not try fallback models, because it is a hard deadline for the whole run. A stepMs timeout is not retried against the same model but does advance to the next entry in models when fallback models are configured. Also accepts firstChunkMs, which only applies to streaming calls and is the maximum time the model may take to emit its first content-bearing chunk (text, reasoning, tool call, file or source; stream-start and metadata chunks do not count). A firstChunkMs timeout fails with timeoutType: 'firstChunk', behaves like stepMs for fallback, and is reset for each provider retry attempt. Nested timeout keys are merged across call-time and per-model settings.
171
+ **options.modelSettings.timeout** (`object`): Time-based execution budget for the run. It must be an object whose configured values are positive, finite numbers of milliseconds. Accepts totalMs, the maximum duration of the entire agent run across every loop iteration, tool call and retry, and stepMs, the maximum duration of a single model call including the time spent consuming its stream. Exceeding either budget fails with a MastraTimeoutError. A totalMs timeout ends the run and does not try fallback models, because it is a hard deadline for the whole run. A stepMs timeout is not retried against the same model but does advance to the next entry in models when fallback models are configured. Also accepts firstChunkMs, which only applies to streaming calls and is the maximum time the model may take to emit its first content-bearing chunk (text, reasoning, tool call, file or source; stream-start and metadata chunks do not count). A firstChunkMs timeout fails with timeoutType: 'firstChunk', behaves like stepMs for fallback, and is reset for each provider retry attempt. Nested timeout keys are merged across call-time and per-model settings.
172
172
 
173
173
  **options.modelSettings.stopSequences** (`string[]`): Stop sequences. If set, the model will stop generating text when one of the stop sequences is generated.
174
174