@mastra/mcp-docs-server 1.2.24-alpha.14 → 1.2.24-alpha.18

This diff represents the content of publicly available package versions that have been released to one of the supported registries. The information contained in this diff is provided for informational purposes only and reflects changes between package versions as they appear in their respective public registries.
@@ -189,6 +189,17 @@ export const durableAgent = createDurableAgent({
189
189
 
190
190
  `createInngestAgent()` doesn't enable caching by default. Pass a `cache` option or register the agent with a `Mastra` instance that has a `serverCache` configured to enable resumable streams.
191
191
 
192
+ For cached topics, each published event is recorded in the cache before it's delivered live, so publish latency is bounded by the round-trip to the cache. If the cache write fails, the event is still delivered live but can't be replayed. `@mastra/redis` and `@mastra/valkey` record each event in a single round-trip using a Lua script. `@mastra/redis` falls back to separate commands when the client has no `evalScript` or when Redis Cluster rejects the multi-key script. Per-run workflow watch events (`workflow.events.v2.*`) are never cached. If the cache is remote (for example, in another region) and you don't need to resume a topic, use `shouldCache` to publish that topic straight through:
193
+
194
+ ```typescript
195
+ export const durableAgent = createDurableAgent({
196
+ agent,
197
+ cache,
198
+ // Skip the replay cache for the per-chunk stream topic; other topics stay resumable.
199
+ shouldCache: topic => !topic.startsWith('agent.stream.'),
200
+ })
201
+ ```
202
+
192
203
  ## Streaming with background tasks
193
204
 
194
205
  Durable agents support the same [`untilIdle`](https://mastra.ai/reference/streaming/agents/stream) option as regular agents. When `untilIdle` is set, `stream()` keeps the connection open across background-task continuations until the agent is idle:
@@ -216,10 +216,13 @@ const tracingOptions = {
216
216
 
217
217
  This example produces `langfuse.trace.metadata.customerId` and `langfuse.trace.metadata.tier`.
218
218
 
219
+ Metadata on the root span is also forwarded. Mastra sets `runId` and `resourceId` on every agent and workflow root span, and you can add your own keys through `tracingOptions.metadata`. The exporter forwards each of these root span keys to `langfuse.trace.metadata.<key>`. Keys that map to a dedicated Langfuse field (`userId`, `sessionId`, `threadId`, `traceName`, and `version`) are not duplicated as trace metadata. Metadata on child spans stays on the observation and never changes the trace.
220
+
219
221
  Notes:
220
222
 
221
223
  - The reserved `prompt` key is used for [prompt linking](#prompt-linking) and isn't forwarded as trace metadata.
222
224
  - The reserved identity keys `agentId`, `agentName`, `workflowId`, and `workflowName` are set from the root span and take precedence over custom values with the same name.
225
+ - Keys under `langfuse` take precedence over root span metadata with the same name.
223
226
  - Values are sent as strings, because Langfuse maps trace metadata attributes as strings. Numbers, booleans, and objects are serialized with JSON. Langfuse Cloud restores them to their original types on ingestion.
224
227
 
225
228
  ## Prompt linking
@@ -4,7 +4,7 @@
4
4
 
5
5
  # ![Neon logo](https://models.dev/logos/neon.svg)Neon
6
6
 
7
- Neon aggregates models from multiple providers with enhanced features like rate limiting and failover. Access 42 models through Mastra's model router.
7
+ Neon aggregates models from multiple providers with enhanced features like rate limiting and failover. Access 46 models through Mastra's model router.
8
8
 
9
9
  Learn more in the [Neon documentation](https://neon.com/docs).
10
10
 
@@ -36,6 +36,7 @@ NEON_AI_GATEWAY_TOKEN=your-gateway-key
36
36
  | Model |
37
37
  | ----------------------------- |
38
38
  | `claude-fable-5` |
39
+ | `claude-fable-5-1` |
39
40
  | `claude-haiku-4-5` |
40
41
  | `claude-opus-4-1` |
41
42
  | `claude-opus-4-5` |
@@ -54,6 +55,7 @@ NEON_AI_GATEWAY_TOKEN=your-gateway-key
54
55
  | `gemini-3-flash` |
55
56
  | `gemma-3-12b` |
56
57
  | `glm-5-2` |
58
+ | `glm-5-3-flash` |
57
59
  | `gpt-5` |
58
60
  | `gpt-5-1` |
59
61
  | `gpt-5-2` |
@@ -68,8 +70,10 @@ NEON_AI_GATEWAY_TOKEN=your-gateway-key
68
70
  | `gpt-5-6-terra` |
69
71
  | `gpt-5-mini` |
70
72
  | `gpt-5-nano` |
73
+ | `gpt-6-astra` |
71
74
  | `gpt-oss-120b` |
72
75
  | `gpt-oss-20b` |
76
+ | `grok-4-6` |
73
77
  | `inkling` |
74
78
  | `kimi-k3` |
75
79
  | `llama-4-maverick` |
@@ -4,7 +4,7 @@
4
4
 
5
5
  # Netlify
6
6
 
7
- Netlify AI Gateway provides unified access to multiple providers with built-in caching and observability. Access 243 models through Mastra's model router.
7
+ Netlify AI Gateway provides unified access to multiple providers with built-in caching and observability. Access 242 models through Mastra's model router.
8
8
 
9
9
  Learn more in the [Netlify documentation](https://docs.netlify.com/build/ai-gateway/overview/).
10
10
 
@@ -279,6 +279,5 @@ ANTHROPIC_API_KEY=ant-...
279
279
  | `openrouter/z-ai/glm-5` |
280
280
  | `openrouter/z-ai/glm-5.1` |
281
281
  | `openrouter/z-ai/glm-5.2` |
282
- | `openrouter/z-ai/glm-5.2:free` |
283
282
  | `openrouter/z-ai/glm-5.3` |
284
283
  | `openrouter/z-ai/glm-5.3-flash` |
@@ -4,7 +4,7 @@
4
4
 
5
5
  # ![OpenRouter logo](https://models.dev/logos/openrouter.svg)OpenRouter
6
6
 
7
- OpenRouter aggregates models from multiple providers with enhanced features like rate limiting and failover. Access 361 models through Mastra's model router.
7
+ OpenRouter aggregates models from multiple providers with enhanced features like rate limiting and failover. Access 358 models through Mastra's model router.
8
8
 
9
9
  Learn more in the [OpenRouter documentation](https://openrouter.ai/models).
10
10
 
@@ -174,9 +174,7 @@ ANTHROPIC_API_KEY=ant-...
174
174
  | `minimax/minimax-m2.1` |
175
175
  | `minimax/minimax-m2.5` |
176
176
  | `minimax/minimax-m2.7` |
177
- | `minimax/minimax-m2.7:free` |
178
177
  | `minimax/minimax-m3` |
179
- | `minimax/minimax-m3:free` |
180
178
  | `mistralai/codestral-2508` |
181
179
  | `mistralai/devstral-2512` |
182
180
  | `mistralai/ministral-14b-2512` |
@@ -395,7 +393,6 @@ ANTHROPIC_API_KEY=ant-...
395
393
  | `z-ai/glm-5-turbo` |
396
394
  | `z-ai/glm-5.1` |
397
395
  | `z-ai/glm-5.2` |
398
- | `z-ai/glm-5.2:free` |
399
396
  | `z-ai/glm-5.3` |
400
397
  | `z-ai/glm-5.3-flash` |
401
398
  | `z-ai/glm-5v-turbo` |
@@ -4,7 +4,7 @@
4
4
 
5
5
  # ![Vercel logo](https://models.dev/logos/vercel.svg)Vercel
6
6
 
7
- Vercel aggregates models from multiple providers with enhanced features like rate limiting and failover. Access 373 models through Mastra's model router.
7
+ Vercel aggregates models from multiple providers with enhanced features like rate limiting and failover. Access 371 models through Mastra's model router.
8
8
 
9
9
  Learn more in the [Vercel documentation](https://ai-sdk.dev/providers/ai-sdk-providers).
10
10
 
@@ -221,10 +221,8 @@ ANTHROPIC_API_KEY=ant-...
221
221
  | `minimax/minimax-m2.5` |
222
222
  | `minimax/minimax-m2.5-highspeed` |
223
223
  | `minimax/minimax-m2.7` |
224
- | `minimax/minimax-m2.7-free` |
225
224
  | `minimax/minimax-m2.7-highspeed` |
226
225
  | `minimax/minimax-m3` |
227
- | `minimax/minimax-m3-free` |
228
226
  | `mistral/codestral` |
229
227
  | `mistral/codestral-embed` |
230
228
  | `mistral/devstral-2` |
@@ -4,7 +4,7 @@
4
4
 
5
5
  # Model Providers
6
6
 
7
- Mastra provides a unified interface for working with LLMs across multiple providers, giving you access to 7106 models from 200 providers through a single API.
7
+ Mastra provides a unified interface for working with LLMs across multiple providers, giving you access to 7098 models from 200 providers through a single API.
8
8
 
9
9
  ## Features
10
10
 
@@ -19,7 +19,7 @@ const agent = new Agent({
19
19
  id: "my-agent",
20
20
  name: "My Agent",
21
21
  instructions: "You are a helpful assistant",
22
- model: "cerebras/gemma-4-31b"
22
+ model: "cerebras/gpt-oss-120b"
23
23
  });
24
24
 
25
25
  // Generate a response
@@ -36,8 +36,8 @@ for await (const chunk of stream) {
36
36
 
37
37
  | Model | Context | Tools | Reasoning | Image | Audio | Video | Input $/1M | Output $/1M |
38
38
  | ----------------------- | ------- | ----- | --------- | ----- | ----- | ----- | ---------- | ----------- |
39
- | `cerebras/gemma-4-31b` | 131K | | | | | | $0.99 | $1 |
40
39
  | `cerebras/gpt-oss-120b` | 131K | | | | | | $0.35 | $0.75 |
40
+ | `cerebras/qwen-3.8-27b` | 66K | | | | | | $0.99 | $1 |
41
41
 
42
42
  Model availability, capabilities, context windows, and pricing are sourced from [models.dev](https://models.dev) and may change.
43
43
 
@@ -50,7 +50,7 @@ const agent = new Agent({
50
50
  id: "custom-agent",
51
51
  name: "custom-agent",
52
52
  model: {
53
- id: "cerebras/gemma-4-31b",
53
+ id: "cerebras/gpt-oss-120b",
54
54
  apiKey: process.env.CEREBRAS_API_KEY,
55
55
  headers: {
56
56
  "X-Custom-Header": "value"
@@ -68,8 +68,8 @@ const agent = new Agent({
68
68
  model: ({ requestContext }) => {
69
69
  const useAdvanced = requestContext.task === "complex";
70
70
  return useAdvanced
71
- ? "cerebras/gpt-oss-120b"
72
- : "cerebras/gemma-4-31b";
71
+ ? "cerebras/qwen-3.8-27b"
72
+ : "cerebras/gpt-oss-120b";
73
73
  }
74
74
  });
75
75
  ```
@@ -4,7 +4,7 @@
4
4
 
5
5
  # ![Cortecs logo](https://models.dev/logos/cortecs.svg)Cortecs
6
6
 
7
- Access 109 Cortecs models through Mastra's model router. Authentication is handled automatically using the `CORTECS_API_KEY` environment variable.
7
+ Access 105 Cortecs models through Mastra's model router. Authentication is handled automatically using the `CORTECS_API_KEY` environment variable.
8
8
 
9
9
  Learn more in the [Cortecs documentation](https://cortecs.ai).
10
10
 
@@ -50,7 +50,6 @@ for await (const chunk of stream) {
50
50
  | `cortecs/claude-sonnet-4` | 200K | | | | | | $3 | $14 |
51
51
  | `cortecs/claude-sonnet-5` | 1.0M | | | | | | $2 | $11 |
52
52
  | `cortecs/codestral-2508` | 256K | | | | | | $0.33 | $1 |
53
- | `cortecs/cosmos3-super-reasoner` | 256K | | | | | | $0.10 | $0.30 |
54
53
  | `cortecs/deepseek-r1-0528` | 164K | | | | | | $0.65 | $3 |
55
54
  | `cortecs/deepseek-v3.2` | 164K | | | | | | $0.30 | $0.49 |
56
55
  | `cortecs/deepseek-v4-flash-0731` | 1.0M | | | | | | $0.13 | $0.28 |
@@ -98,7 +97,6 @@ for await (const chunk of stream) {
98
97
  | `cortecs/kimi-k3` | 1.0M | | | | | | $3 | $15 |
99
98
  | `cortecs/llama-3.1-405b-instruct` | 128K | | | | | | $2 | $2 |
100
99
  | `cortecs/llama-3.1-8b-instruct` | 128K | | | | | | $0.17 | $0.17 |
101
- | `cortecs/llama-3.1-nemotron-ultra-253b-v1` | 128K | | | | | | $0.60 | $2 |
102
100
  | `cortecs/llama-3.3-70b-instruct` | 131K | | | | | | $0.13 | $0.40 |
103
101
  | `cortecs/minicpm-v-4.5` | 32K | | | | | | $0.65 | $1 |
104
102
  | `cortecs/minimax-m2` | 400K | | | | | | $0.35 | $1 |
@@ -125,7 +123,6 @@ for await (const chunk of stream) {
125
123
  | `cortecs/nova-micro-v1` | 128K | | | | | | $0.04 | $0.16 |
126
124
  | `cortecs/nova-pro-v1` | 300K | | | | | | $0.92 | $4 |
127
125
  | `cortecs/nvidia-nemotron-3-nano-30b-a3b` | 256K | | | | | | $0.06 | $0.24 |
128
- | `cortecs/nvidia-nemotron-3-nano-omni` | 300K | | | | | | $0.06 | $0.24 |
129
126
  | `cortecs/pixtral-12b-2409` | 128K | | | | | | $0.22 | $0.22 |
130
127
  | `cortecs/pixtral-large-2502` | 128K | | | | | | $2 | $6 |
131
128
  | `cortecs/qwen2.5-vl-72b-instruct` | 32K | | | | | | $1 | $1 |
@@ -134,7 +131,6 @@ for await (const chunk of stream) {
134
131
  | `cortecs/qwen3-32b` | 32K | | | | | | $0.09 | $0.31 |
135
132
  | `cortecs/qwen3-coder-30b-a3b-instruct` | 262K | | | | | | $0.07 | $0.24 |
136
133
  | `cortecs/qwen3-coder-next` | 256K | | | | | | $0.17 | $0.89 |
137
- | `cortecs/qwen3-next-80b-a3b-thinking` | 128K | | | | | | $0.15 | $1 |
138
134
  | `cortecs/qwen3-vl-235b-a22b` | 256K | | | | | | $0.62 | $3 |
139
135
  | `cortecs/qwen3.5-122b-a10b` | 262K | | | | | | $0.49 | $3 |
140
136
  | `cortecs/qwen3.5-397b-a17b` | 262K | | | | | | $0.67 | $4 |
@@ -4,7 +4,7 @@
4
4
 
5
5
  # ![CrofAI logo](https://models.dev/logos/crof.svg)CrofAI
6
6
 
7
- Access 23 CrofAI models through Mastra's model router. Authentication is handled automatically using the `CROF_API_KEY` environment variable.
7
+ Access 24 CrofAI models through Mastra's model router. Authentication is handled automatically using the `CROF_API_KEY` environment variable.
8
8
 
9
9
  Learn more in the [CrofAI documentation](https://crof.ai/docs).
10
10
 
@@ -36,31 +36,32 @@ for await (const chunk of stream) {
36
36
 
37
37
  ## Models
38
38
 
39
- | Model | Context | Tools | Reasoning | Image | Audio | Video | Input $/1M | Output $/1M |
40
- | ----------------------------- | ------- | ----- | --------- | ----- | ----- | ----- | ---------- | ----------- |
41
- | `crof/deepseek-v3.2` | 164K | | | | | | $0.18 | $0.35 |
42
- | `crof/deepseek-v4-flash` | 1.0M | | | | | | $0.12 | $0.21 |
43
- | `crof/deepseek-v4-flash-0731` | 1.0M | | | | | | $0.08 | $0.10 |
44
- | `crof/deepseek-v4-pro` | 1.0M | | | | | | $0.35 | $0.80 |
45
- | `crof/deepseek-v4-pro-0813` | 1.0M | | | | | | $0.35 | $0.80 |
46
- | `crof/gemma-4-31b-it` | 262K | | | | | | $0.10 | $0.30 |
47
- | `crof/glm-5.1` | 203K | | | | | | $0.45 | $2 |
48
- | `crof/glm-5.2` | 1.0M | | | | | | $0.30 | $1 |
49
- | `crof/glm-5.3` | 1.0M | | | | | | $0.40 | $1 |
50
- | `crof/glm-5.3-flash` | 1.0M | | | | | | $0.07 | $0.22 |
51
- | `crof/greg-1-mini` | 229K | | | | | | $0.07 | $0.15 |
52
- | `crof/greg-2-super` | 229K | | | | | | $2 | $5 |
53
- | `crof/greg-2-ultra` | 229K | | | | | | $3 | $10 |
54
- | `crof/greg-rp` | 229K | | | | | | $0.10 | $0.30 |
55
- | `crof/kimi-k2.6` | 262K | | | | | | $0.50 | $2 |
56
- | `crof/kimi-k2.7-code` | 262K | | | | | | $0.55 | $2 |
57
- | `crof/kimi-k3` | 1.0M | | | | | | $2 | $8 |
58
- | `crof/kimi-k3-eco` | 1.0M | | | | | | $1 | $4 |
59
- | `crof/mimo-v2.5-pro` | 1.0M | | | | | | $0.40 | $0.80 |
60
- | `crof/qwen3.5-397b-a17b` | 262K | | | | | | $0.35 | $2 |
61
- | `crof/qwen3.5-9b` | 262K | | | | | | $0.04 | $0.15 |
62
- | `crof/qwen3.6-27b` | 262K | | | | | | $0.20 | $2 |
63
- | `crof/qwen3.8-27b` | 262K | | | | | | $0.20 | $2 |
39
+ | Model | Context | Tools | Reasoning | Image | Audio | Video | Input $/1M | Output $/1M |
40
+ | ----------------------------------- | ------- | ----- | --------- | ----- | ----- | ----- | ---------- | ----------- |
41
+ | `crof/deepseek-v3.2` | 164K | | | | | | $0.18 | $0.35 |
42
+ | `crof/deepseek-v4-flash` | 1.0M | | | | | | $0.12 | $0.21 |
43
+ | `crof/deepseek-v4-flash-0731` | 1.0M | | | | | | $0.08 | $0.10 |
44
+ | `crof/deepseek-v4-flash-vision-exp` | 1.0M | | | | | | $0.08 | $0.20 |
45
+ | `crof/deepseek-v4-pro` | 1.0M | | | | | | $0.35 | $0.80 |
46
+ | `crof/deepseek-v4-pro-0813` | 1.0M | | | | | | $0.35 | $0.80 |
47
+ | `crof/gemma-4-31b-it` | 262K | | | | | | $0.10 | $0.30 |
48
+ | `crof/glm-5.1` | 203K | | | | | | $0.45 | $2 |
49
+ | `crof/glm-5.2` | 1.0M | | | | | | $0.30 | $1 |
50
+ | `crof/glm-5.3` | 1.0M | | | | | | $0.40 | $1 |
51
+ | `crof/glm-5.3-flash` | 1.0M | | | | | | $0.07 | $0.22 |
52
+ | `crof/greg-1-mini` | 229K | | | | | | $0.07 | $0.15 |
53
+ | `crof/greg-2-super` | 229K | | | | | | $2 | $5 |
54
+ | `crof/greg-2-ultra` | 229K | | | | | | $3 | $10 |
55
+ | `crof/greg-rp` | 229K | | | | | | $0.10 | $0.30 |
56
+ | `crof/kimi-k2.6` | 262K | | | | | | $0.50 | $2 |
57
+ | `crof/kimi-k2.7-code` | 262K | | | | | | $0.55 | $2 |
58
+ | `crof/kimi-k3` | 1.0M | | | | | | $2 | $8 |
59
+ | `crof/kimi-k3-eco` | 1.0M | | | | | | $1 | $4 |
60
+ | `crof/mimo-v2.5-pro` | 1.0M | | | | | | $0.40 | $0.80 |
61
+ | `crof/qwen3.5-397b-a17b` | 262K | | | | | | $0.35 | $2 |
62
+ | `crof/qwen3.5-9b` | 262K | | | | | | $0.04 | $0.15 |
63
+ | `crof/qwen3.6-27b` | 262K | | | | | | $0.20 | $2 |
64
+ | `crof/qwen3.8-27b` | 262K | | | | | | $0.20 | $2 |
64
65
 
65
66
  Model availability, capabilities, context windows, and pricing are sourced from [models.dev](https://models.dev) and may change.
66
67
 
@@ -45,7 +45,7 @@ for await (const chunk of stream) {
45
45
  | `deepinfra/deepseek-ai/DeepSeek-V3.1` | 164K | | | | | | $0.25 | $0.95 |
46
46
  | `deepinfra/deepseek-ai/DeepSeek-V3.2` | 164K | | | | | | $0.26 | $0.38 |
47
47
  | `deepinfra/deepseek-ai/DeepSeek-V4-Flash` | 1.0M | | | | | | $0.09 | $0.18 |
48
- | `deepinfra/deepseek-ai/DeepSeek-V4-Flash-0731` | 1.0M | | | | | | $0.08 | $0.18 |
48
+ | `deepinfra/deepseek-ai/DeepSeek-V4-Flash-0731` | 1.0M | | | | | | $0.06 | $0.18 |
49
49
  | `deepinfra/deepseek-ai/DeepSeek-V4-Flash-Vision-Exp` | 1.0M | | | | | | $0.44 | $1 |
50
50
  | `deepinfra/deepseek-ai/DeepSeek-V4-Pro` | 1.0M | | | | | | $1 | $3 |
51
51
  | `deepinfra/deepseek-ai/DeepSeek-V4-Pro-0813` | 1.0M | | | | | | $1 | $3 |
@@ -4,7 +4,7 @@
4
4
 
5
5
  # ![Eden AI logo](https://models.dev/logos/edenai.svg)Eden AI
6
6
 
7
- Access 249 Eden AI models through Mastra's model router. Authentication is handled automatically using the `EDENAI_API_KEY` environment variable.
7
+ Access 250 Eden AI models through Mastra's model router. Authentication is handled automatically using the `EDENAI_API_KEY` environment variable.
8
8
 
9
9
  Learn more in the [Eden AI documentation](https://docs.edenai.co).
10
10
 
@@ -78,12 +78,13 @@ for await (const chunk of stream) {
78
78
  | `edenai/databricks/databricks-gpt-oss-120b@eu` | 131K | | | | | | $0.15 | $0.60 |
79
79
  | `edenai/databricks/databricks-gpt-oss-20b` | 131K | | | | | | $0.07 | $0.30 |
80
80
  | `edenai/databricks/databricks-gpt-oss-20b@eu` | 131K | | | | | | $0.07 | $0.30 |
81
+ | `edenai/databricks/databricks-inkling` | 1.0M | | | | | | $1 | $4 |
81
82
  | `edenai/deepinfra/ByteDance/Seed-2.0-code` | 256K | | | | | | $0.50 | $3 |
82
83
  | `edenai/deepinfra/ByteDance/Seed-2.0-mini` | 256K | | | | | | $0.10 | $0.40 |
83
84
  | `edenai/deepinfra/deepseek-ai/DeepSeek-R1` | 164K | | | | | | $0.70 | $2 |
84
85
  | `edenai/deepinfra/deepseek-ai/DeepSeek-V3` | 164K | | | | | | $0.32 | $0.89 |
85
86
  | `edenai/deepinfra/deepseek-ai/DeepSeek-V3-0324` | 164K | | | | | | $0.24 | $0.90 |
86
- | `edenai/deepinfra/deepseek-ai/DeepSeek-V4-Flash-0731` | 1.0M | | | | | | $0.08 | $0.18 |
87
+ | `edenai/deepinfra/deepseek-ai/DeepSeek-V4-Flash-0731` | 1.0M | | | | | | $0.06 | $0.18 |
87
88
  | `edenai/deepinfra/deepseek-ai/DeepSeek-V4-Pro-0813` | 1.0M | | | | | | $1 | $3 |
88
89
  | `edenai/deepinfra/meta-llama/Llama-3.2-11B-Vision-Instruct` | 131K | | | | | | $0.34 | $0.34 |
89
90
  | `edenai/deepinfra/meta-llama/Llama-3.3-70B-Instruct` | 131K | | | | | | $0.10 | $0.32 |
@@ -110,7 +111,7 @@ for await (const chunk of stream) {
110
111
  | `edenai/fireworks_ai/accounts/fireworks/models/inkling` | 1.0M | | | | | | $1 | $4 |
111
112
  | `edenai/fireworks_ai/accounts/fireworks/models/muse-glimmer-30b` | 131K | | | | | | $0.35 | $2 |
112
113
  | `edenai/fireworks_ai/gpt-oss-120b` | 131K | | | | | | $0.15 | $0.60 |
113
- | `edenai/flexai/deepseek-v4-flash-0731` | 786K | | | | | | $0.03 | $0.10 |
114
+ | `edenai/flexai/DeepSeek-V4-Flash-0731` | 786K | | | | | | $0.03 | $0.10 |
114
115
  | `edenai/flexai/gpt-oss-120b` | 131K | | | | | | $0.04 | $0.10 |
115
116
  | `edenai/flexai/gpt-oss-20b` | 131K | | | | | | $0.02 | $0.10 |
116
117
  | `edenai/flexai/Muse-Glimmer-30B` | 131K | | | | | | $0.30 | $1 |
@@ -146,7 +147,7 @@ for await (const chunk of stream) {
146
147
  | `edenai/mistral/codestral-latest` | 256K | | | | | | $0.30 | $0.90 |
147
148
  | `edenai/mistral/devstral-2512` | 262K | | | | | | $0.40 | $2 |
148
149
  | `edenai/mistral/devstral-medium-latest` | 262K | | | | | | $0.40 | $2 |
149
- | `edenai/mistral/magistral-medium-latest` | 262K | | | | | | $2 | $5 |
150
+ | `edenai/mistral/magistral-medium-latest` | 262K | | | | | | $2 | $8 |
150
151
  | `edenai/mistral/mistral-large-2512` | 262K | | | | | | $0.50 | $2 |
151
152
  | `edenai/mistral/mistral-large-latest` | 262K | | | | | | $2 | $6 |
152
153
  | `edenai/mistral/mistral-medium-2505` | 131K | | | | | | $0.40 | $2 |
@@ -4,7 +4,7 @@
4
4
 
5
5
  # ![Fireworks AI logo](https://models.dev/logos/fireworks-ai.svg)Fireworks AI
6
6
 
7
- Access 20 Fireworks AI models through Mastra's model router. Authentication is handled automatically using the `FIREWORKS_API_KEY` environment variable.
7
+ Access 21 Fireworks AI models through Mastra's model router. Authentication is handled automatically using the `FIREWORKS_API_KEY` environment variable.
8
8
 
9
9
  Learn more in the [Fireworks AI documentation](https://fireworks.ai/docs/).
10
10
 
@@ -57,6 +57,7 @@ for await (const chunk of stream) {
57
57
  | `fireworks-ai/accounts/fireworks/models/qwen3p8-2p4t-a95b` | 262K | | | | | | $2 | $6 |
58
58
  | `fireworks-ai/accounts/fireworks/models/qwen3p8-max` | 262K | | | | | | $2 | $6 |
59
59
  | `fireworks-ai/accounts/fireworks/routers/glm-5p2-fast` | 1.0M | | | | | | $2 | $7 |
60
+ | `fireworks-ai/accounts/fireworks/routers/glm-5p3-fast` | 1.0M | | | | | | $2 | $7 |
60
61
  | `fireworks-ai/accounts/fireworks/routers/kimi-k3-fast` | 1.0M | | | | | | $5 | $23 |
61
62
 
62
63
  Model availability, capabilities, context windows, and pricing are sourced from [models.dev](https://models.dev) and may change.
@@ -42,7 +42,7 @@ for await (const chunk of stream) {
42
42
  | `hyper/deepseek-v4-flash-0731` | 1.0M | | | | | | $0.44 | $1 |
43
43
  | `hyper/deepseek-v4-pro` | 1.0M | | | | | | $2 | $5 |
44
44
  | `hyper/deepseek-v4-pro-0813` | 1.0M | | | | | | $1 | $4 |
45
- | `hyper/gemma-4-26b-a4b-it` | 256K | | | | | | $0.11 | $0.37 |
45
+ | `hyper/gemma-4-26b-a4b-it` | 256K | | | | | | $0.11 | $0.41 |
46
46
  | `hyper/glm-5` | 203K | | | | | | $0.86 | $3 |
47
47
  | `hyper/glm-5.1` | 203K | | | | | | $1 | $4 |
48
48
  | `hyper/glm-5.2` | 1.0M | | | | | | $2 | $5 |
@@ -51,13 +51,13 @@ for await (const chunk of stream) {
51
51
  | `hyper/gpt-oss-120b` | 128K | | | | | | $0.19 | $0.70 |
52
52
  | `hyper/inkling` | 1.0M | | | | | | $1 | $4 |
53
53
  | `hyper/kimi-k2-thinking` | 262K | | | | | | $0.60 | $3 |
54
- | `hyper/kimi-k2.5` | 262K | | | | | | $0.51 | $3 |
54
+ | `hyper/kimi-k2.5` | 262K | | | | | | $0.56 | $3 |
55
55
  | `hyper/kimi-k2.6` | 262K | | | | | | $1 | $4 |
56
56
  | `hyper/kimi-k2.7-code` | 262K | | | | | | $1 | $4 |
57
57
  | `hyper/kimi-k3` | 1.0M | | | | | | $3 | $16 |
58
58
  | `hyper/llama-3.3-70b-instruct` | 128K | | | | | | $0.61 | $1 |
59
59
  | `hyper/llama-4-maverick-17b-128e-instruct-fp8` | 430K | | | | | | $0.27 | $0.90 |
60
- | `hyper/minimax-m2.7` | 262K | | | | | | $0.48 | $2 |
60
+ | `hyper/minimax-m2.7` | 262K | | | | | | $0.47 | $2 |
61
61
  | `hyper/minimax-m3` | 512K | | | | | | $0.33 | $1 |
62
62
  | `hyper/qwen3-coder-480b-a35b-instruct-int4-mixed-ar` | 106K | | | | | | $0.45 | $2 |
63
63
  | `hyper/qwen3-next-80b-a3b-instruct` | 262K | | | | | | $0.12 | $1 |
@@ -4,7 +4,7 @@
4
4
 
5
5
  # ![Kilo Gateway logo](https://models.dev/logos/kilo.svg)Kilo Gateway
6
6
 
7
- Access 369 Kilo Gateway models through Mastra's model router. Authentication is handled automatically using the `KILO_API_KEY` environment variable.
7
+ Access 367 Kilo Gateway models through Mastra's model router. Authentication is handled automatically using the `KILO_API_KEY` environment variable.
8
8
 
9
9
  Learn more in the [Kilo Gateway documentation](https://kilo.ai).
10
10
 
@@ -42,15 +42,15 @@ for await (const chunk of stream) {
42
42
  | `kilo/~anthropic/claude-haiku-latest` | 200K | | | | | | $1 | $5 |
43
43
  | `kilo/~anthropic/claude-opus-latest` | 1.0M | | | | | | $5 | $25 |
44
44
  | `kilo/~anthropic/claude-sonnet-latest` | 1.0M | | | | | | $2 | $10 |
45
- | `kilo/~deepseek/deepseek-v4-flash-latest` | 1.0M | | | | | | $0.05 | $0.10 |
45
+ | `kilo/~deepseek/deepseek-v4-flash-latest` | 1.0M | | | | | | $0.05 | $0.16 |
46
46
  | `kilo/~google/gemini-flash-latest` | 1.0M | | | | | | $0.75 | $4 |
47
47
  | `kilo/~google/gemini-pro-latest` | 1.0M | | | | | | $2 | $12 |
48
- | `kilo/~moonshotai/kimi-latest` | 1.0M | | | | | | $3 | $13 |
48
+ | `kilo/~moonshotai/kimi-latest` | 1.0M | | | | | | $3 | $14 |
49
49
  | `kilo/~openai/gpt-latest` | 1.1M | | | | | | $2 | $10 |
50
50
  | `kilo/~openai/gpt-mini-latest` | 400K | | | | | | $0.75 | $5 |
51
51
  | `kilo/~x-ai/grok-latest` | 500K | | | | | | $2 | $6 |
52
- | `kilo/~z-ai/glm-flash-latest` | 1.0M | | | | | | $0.07 | $0.25 |
53
- | `kilo/~z-ai/glm-latest` | 262K | | | | | | $1 | $4 |
52
+ | `kilo/~z-ai/glm-flash-latest` | 1.0M | | | | | | $0.07 | $0.24 |
53
+ | `kilo/~z-ai/glm-latest` | 1.0M | | | | | | $1 | $4 |
54
54
  | `kilo/aion-labs/aion-2.0` | 131K | | | | | | $0.80 | $2 |
55
55
  | `kilo/aion-labs/aion-3.0` | 131K | | | | | | $3 | $6 |
56
56
  | `kilo/aion-labs/aion-3.0-mini` | 131K | | | | | | $0.70 | $1 |
@@ -177,9 +177,7 @@ for await (const chunk of stream) {
177
177
  | `kilo/minimax/minimax-m2.1` | 205K | | | | | | $0.30 | $1 |
178
178
  | `kilo/minimax/minimax-m2.5` | 200K | | | | | | $0.30 | $1 |
179
179
  | `kilo/minimax/minimax-m2.7` | 205K | | | | | | $0.30 | $1 |
180
- | `kilo/minimax/minimax-m2.7:free` | 197K | | | | | | — | — |
181
180
  | `kilo/minimax/minimax-m3` | 524K | | | | | | $0.30 | $1 |
182
- | `kilo/minimax/minimax-m3:free` | 1.0M | | | | | | — | — |
183
181
  | `kilo/mistralai/codestral-2508` | 256K | | | | | | $0.30 | $0.90 |
184
182
  | `kilo/mistralai/devstral-2512` | 262K | | | | | | $0.40 | $2 |
185
183
  | `kilo/mistralai/ministral-14b-2512` | 262K | | | | | | $0.20 | $0.20 |
@@ -265,7 +263,7 @@ for await (const chunk of stream) {
265
263
  | `kilo/openai/gpt-5.6-luna` | 1.1M | | | | | | $0.20 | $1 |
266
264
  | `kilo/openai/gpt-5.6-luna-pro` | 1.1M | | | | | | $0.20 | $1 |
267
265
  | `kilo/openai/gpt-5.6-sol` | 1.1M | | | | | | $4 | $20 |
268
- | `kilo/openai/gpt-5.6-sol-discounted` | 1.1M | | | | | | $3 | $15 |
266
+ | `kilo/openai/gpt-5.6-sol-discounted` | 1.1M | | | | | | $2 | $10 |
269
267
  | `kilo/openai/gpt-5.6-sol-pro` | 1.1M | | | | | | $4 | $20 |
270
268
  | `kilo/openai/gpt-5.6-terra` | 1.1M | | | | | | $2 | $12 |
271
269
  | `kilo/openai/gpt-5.6-terra-pro` | 1.1M | | | | | | $2 | $12 |
@@ -322,7 +320,7 @@ for await (const chunk of stream) {
322
320
  | `kilo/qwen/qwen3-max` | 262K | | | | | | $0.78 | $4 |
323
321
  | `kilo/qwen/qwen3-max-thinking` | 262K | | | | | | $0.78 | $4 |
324
322
  | `kilo/qwen/qwen3-next-80b-a3b-instruct` | 262K | | | | | | $0.10 | $0.78 |
325
- | `kilo/qwen/qwen3-next-80b-a3b-thinking` | 131K | | | | | | $0.15 | $1 |
323
+ | `kilo/qwen/qwen3-next-80b-a3b-thinking` | 262K | | | | | | $0.15 | $1 |
326
324
  | `kilo/qwen/qwen3-vl-235b-a22b-instruct` | 131K | | | | | | $0.26 | $1 |
327
325
  | `kilo/qwen/qwen3-vl-235b-a22b-thinking` | 131K | | | | | | $0.40 | $4 |
328
326
  | `kilo/qwen/qwen3-vl-30b-a3b-instruct` | 262K | | | | | | $0.13 | $0.52 |
@@ -396,7 +394,7 @@ for await (const chunk of stream) {
396
394
  | `kilo/z-ai/glm-4.5` | 131K | | | | | | $0.60 | $2 |
397
395
  | `kilo/z-ai/glm-4.5-air` | 131K | | | | | | $0.13 | $0.85 |
398
396
  | `kilo/z-ai/glm-4.5v` | 66K | | | | | | $0.60 | $2 |
399
- | `kilo/z-ai/glm-4.6` | 205K | | | | | | $0.55 | $2 |
397
+ | `kilo/z-ai/glm-4.6` | 198K | | | | | | $0.43 | $2 |
400
398
  | `kilo/z-ai/glm-4.6v` | 131K | | | | | | $0.30 | $0.90 |
401
399
  | `kilo/z-ai/glm-4.7` | 203K | | | | | | $0.40 | $2 |
402
400
  | `kilo/z-ai/glm-4.7-flash` | 203K | | | | | | $0.06 | $0.40 |
@@ -4,7 +4,7 @@
4
4
 
5
5
  # ![LLM Gateway logo](https://models.dev/logos/llmgateway-providers.svg)LLM Gateway
6
6
 
7
- Access 374 LLM Gateway models through Mastra's model router. Authentication is handled automatically using the `LLMGATEWAY_API_KEY` environment variable.
7
+ Access 373 LLM Gateway models through Mastra's model router. Authentication is handled automatically using the `LLMGATEWAY_API_KEY` environment variable.
8
8
 
9
9
  Learn more in the [LLM Gateway documentation](https://llmgateway.io/docs).
10
10
 
@@ -176,7 +176,6 @@ for await (const chunk of stream) {
176
176
  | `llmgateway-providers/deepinfra/gemma-4-31b-it` | 262K | | | | | | $0.13 | $0.38 |
177
177
  | `llmgateway-providers/deepinfra/glm-5.1` | 198K | | | | | | $1 | $4 |
178
178
  | `llmgateway-providers/deepinfra/hy3` | 262K | | | | | | $0.14 | $0.58 |
179
- | `llmgateway-providers/deepinfra/kimi-k2.5` | 256K | | | | | | $0.45 | $2 |
180
179
  | `llmgateway-providers/deepinfra/ling-3.0-flash` | 262K | | | | | | $0.06 | $0.18 |
181
180
  | `llmgateway-providers/deepinfra/mimo-v2.5` | 262K | | | | | | $0.40 | $2 |
182
181
  | `llmgateway-providers/deepinfra/mimo-v2.5-pro` | 1.0M | | | | | | $1 | $3 |
@@ -4,7 +4,7 @@
4
4
 
5
5
  # ![Moonshot AI (China) logo](https://models.dev/logos/moonshotai-cn.svg)Moonshot AI (China)
6
6
 
7
- Access 10 Moonshot AI (China) models through Mastra's model router. Authentication is handled automatically using the `MOONSHOT_API_KEY` environment variable.
7
+ Access 4 Moonshot AI (China) models through Mastra's model router. Authentication is handled automatically using the `MOONSHOT_API_KEY` environment variable.
8
8
 
9
9
  Learn more in the [Moonshot AI (China) documentation](https://platform.moonshot.cn).
10
10
 
@@ -19,7 +19,7 @@ const agent = new Agent({
19
19
  id: "my-agent",
20
20
  name: "My Agent",
21
21
  instructions: "You are a helpful assistant",
22
- model: "moonshotai-cn/kimi-k2-0711-preview"
22
+ model: "moonshotai-cn/kimi-k2.6"
23
23
  });
24
24
 
25
25
  // Generate a response
@@ -38,12 +38,6 @@ for await (const chunk of stream) {
38
38
 
39
39
  | Model | Context | Tools | Reasoning | Image | Audio | Video | Input $/1M | Output $/1M |
40
40
  | ---------------------------------------- | ------- | ----- | --------- | ----- | ----- | ----- | ---------- | ----------- |
41
- | `moonshotai-cn/kimi-k2-0711-preview` | 131K | | | | | | $0.60 | $3 |
42
- | `moonshotai-cn/kimi-k2-0905-preview` | 262K | | | | | | $0.60 | $3 |
43
- | `moonshotai-cn/kimi-k2-thinking` | 262K | | | | | | $0.60 | $3 |
44
- | `moonshotai-cn/kimi-k2-thinking-turbo` | 262K | | | | | | $1 | $8 |
45
- | `moonshotai-cn/kimi-k2-turbo-preview` | 262K | | | | | | $2 | $10 |
46
- | `moonshotai-cn/kimi-k2.5` | 262K | | | | | | $0.60 | $3 |
47
41
  | `moonshotai-cn/kimi-k2.6` | 262K | | | | | | $0.95 | $4 |
48
42
  | `moonshotai-cn/kimi-k2.7-code` | 262K | | | | | | $0.95 | $4 |
49
43
  | `moonshotai-cn/kimi-k2.7-code-highspeed` | 262K | | | | | | $2 | $8 |
@@ -61,7 +55,7 @@ const agent = new Agent({
61
55
  name: "custom-agent",
62
56
  model: {
63
57
  url: "https://api.moonshot.cn/anthropic/v1",
64
- id: "moonshotai-cn/kimi-k2-0711-preview",
58
+ id: "moonshotai-cn/kimi-k2.6",
65
59
  apiKey: process.env.MOONSHOT_API_KEY,
66
60
  headers: {
67
61
  "X-Custom-Header": "value"
@@ -80,7 +74,7 @@ const agent = new Agent({
80
74
  const useAdvanced = requestContext.task === "complex";
81
75
  return useAdvanced
82
76
  ? "moonshotai-cn/kimi-k3"
83
- : "moonshotai-cn/kimi-k2-0711-preview";
77
+ : "moonshotai-cn/kimi-k2.6";
84
78
  }
85
79
  });
86
80
  ```
@@ -4,7 +4,7 @@
4
4
 
5
5
  # ![Moonshot AI logo](https://models.dev/logos/moonshotai.svg)Moonshot AI
6
6
 
7
- Access 10 Moonshot AI models through Mastra's model router. Authentication is handled automatically using the `MOONSHOT_API_KEY` environment variable.
7
+ Access 4 Moonshot AI models through Mastra's model router. Authentication is handled automatically using the `MOONSHOT_API_KEY` environment variable.
8
8
 
9
9
  Learn more in the [Moonshot AI documentation](https://platform.moonshot.ai).
10
10
 
@@ -19,7 +19,7 @@ const agent = new Agent({
19
19
  id: "my-agent",
20
20
  name: "My Agent",
21
21
  instructions: "You are a helpful assistant",
22
- model: "moonshotai/kimi-k2-0711-preview"
22
+ model: "moonshotai/kimi-k2.6"
23
23
  });
24
24
 
25
25
  // Generate a response
@@ -38,12 +38,6 @@ for await (const chunk of stream) {
38
38
 
39
39
  | Model | Context | Tools | Reasoning | Image | Audio | Video | Input $/1M | Output $/1M |
40
40
  | ------------------------------------- | ------- | ----- | --------- | ----- | ----- | ----- | ---------- | ----------- |
41
- | `moonshotai/kimi-k2-0711-preview` | 131K | | | | | | $0.60 | $3 |
42
- | `moonshotai/kimi-k2-0905-preview` | 262K | | | | | | $0.60 | $3 |
43
- | `moonshotai/kimi-k2-thinking` | 262K | | | | | | $0.60 | $3 |
44
- | `moonshotai/kimi-k2-thinking-turbo` | 262K | | | | | | $1 | $8 |
45
- | `moonshotai/kimi-k2-turbo-preview` | 262K | | | | | | $2 | $10 |
46
- | `moonshotai/kimi-k2.5` | 262K | | | | | | $0.60 | $3 |
47
41
  | `moonshotai/kimi-k2.6` | 262K | | | | | | $0.95 | $4 |
48
42
  | `moonshotai/kimi-k2.7-code` | 262K | | | | | | $0.95 | $4 |
49
43
  | `moonshotai/kimi-k2.7-code-highspeed` | 262K | | | | | | $2 | $8 |
@@ -61,7 +55,7 @@ const agent = new Agent({
61
55
  name: "custom-agent",
62
56
  model: {
63
57
  url: "https://api.moonshot.ai/anthropic/v1",
64
- id: "moonshotai/kimi-k2-0711-preview",
58
+ id: "moonshotai/kimi-k2.6",
65
59
  apiKey: process.env.MOONSHOT_API_KEY,
66
60
  headers: {
67
61
  "X-Custom-Header": "value"
@@ -80,7 +74,7 @@ const agent = new Agent({
80
74
  const useAdvanced = requestContext.task === "complex";
81
75
  return useAdvanced
82
76
  ? "moonshotai/kimi-k3"
83
- : "moonshotai/kimi-k2-0711-preview";
77
+ : "moonshotai/kimi-k2.6";
84
78
  }
85
79
  });
86
80
  ```
@@ -4,7 +4,7 @@
4
4
 
5
5
  # ![NanoGPT logo](https://models.dev/logos/nano-gpt.svg)NanoGPT
6
6
 
7
- Access 591 NanoGPT models through Mastra's model router. Authentication is handled automatically using the `NANO_GPT_API_KEY` environment variable.
7
+ Access 593 NanoGPT models through Mastra's model router. Authentication is handled automatically using the `NANO_GPT_API_KEY` environment variable.
8
8
 
9
9
  Learn more in the [NanoGPT documentation](https://docs.nano-gpt.com).
10
10
 
@@ -152,6 +152,7 @@ for await (const chunk of stream) {
152
152
  | `nano-gpt/deepseek/deepseek-v4-flash-0731:thinking` | 1.0M | | | | | | $0.14 | $0.28 |
153
153
  | `nano-gpt/deepseek/deepseek-v4-flash-latest` | 1.0M | | | | | | $0.14 | $0.28 |
154
154
  | `nano-gpt/deepseek/deepseek-v4-flash-vision-exp` | 1.0M | | | | | | $0.22 | $0.66 |
155
+ | `nano-gpt/deepseek/deepseek-v4-flash-vision-exp-uncensored` | 524K | | | | | | $0.44 | $2 |
155
156
  | `nano-gpt/deepseek/deepseek-v4-flash:thinking` | 1.0M | | | | | | $0.14 | $0.28 |
156
157
  | `nano-gpt/deepseek/deepseek-v4-pro` | 1.0M | | | | | | $1 | $2 |
157
158
  | `nano-gpt/deepseek/deepseek-v4-pro-0813` | 1.0M | | | | | | $1 | $3 |
@@ -381,17 +382,17 @@ for await (const chunk of stream) {
381
382
  | `nano-gpt/openai/gpt-5-nano` | 400K | | | | | | $0.05 | $0.40 |
382
383
  | `nano-gpt/openai/gpt-5-pro` | 400K | | | | | | $15 | $120 |
383
384
  | `nano-gpt/openai/gpt-5.1` | 400K | | | | | | $1 | $10 |
384
- | `nano-gpt/openai/gpt-5.1-2025-11-13` | 1.0M | | | | | | $1 | $10 |
385
+ | `nano-gpt/openai/gpt-5.1-2025-11-13` | 400K | | | | | | $1 | $10 |
385
386
  | `nano-gpt/openai/gpt-5.1-codex` | 400K | | | | | | $1 | $10 |
386
387
  | `nano-gpt/openai/gpt-5.1-codex-max` | 400K | | | | | | $3 | $20 |
387
388
  | `nano-gpt/openai/gpt-5.1-codex-mini` | 400K | | | | | | $0.25 | $2 |
388
389
  | `nano-gpt/openai/gpt-5.2` | 400K | | | | | | $2 | $14 |
389
390
  | `nano-gpt/openai/gpt-5.2-codex` | 400K | | | | | | $2 | $14 |
390
391
  | `nano-gpt/openai/gpt-5.3-codex` | 400K | | | | | | $2 | $14 |
391
- | `nano-gpt/openai/gpt-5.4` | 922K | | | | | | $3 | $15 |
392
+ | `nano-gpt/openai/gpt-5.4` | 1.1M | | | | | | $3 | $15 |
392
393
  | `nano-gpt/openai/gpt-5.4-mini` | 400K | | | | | | $0.75 | $5 |
393
394
  | `nano-gpt/openai/gpt-5.4-nano` | 400K | | | | | | $0.20 | $1 |
394
- | `nano-gpt/openai/gpt-5.5` | 1.0M | | | | | | $5 | $30 |
395
+ | `nano-gpt/openai/gpt-5.5` | 1.1M | | | | | | $5 | $30 |
395
396
  | `nano-gpt/openai/gpt-5.6-luna` | 1.1M | | | | | | $0.20 | $1 |
396
397
  | `nano-gpt/openai/gpt-5.6-luna-pro` | 1.1M | | | | | | $0.20 | $1 |
397
398
  | `nano-gpt/openai/gpt-5.6-sol` | 1.1M | | | | | | $2 | $10 |
@@ -426,6 +427,7 @@ for await (const chunk of stream) {
426
427
  | `nano-gpt/pokee-isaac` | 10.0M | | | | | | $0.15 | $1 |
427
428
  | `nano-gpt/poolside/laguna-s-2.1` | 1.0M | | | | | | $0.10 | $0.20 |
428
429
  | `nano-gpt/poolside/laguna-s-2.1:thinking` | 1.0M | | | | | | $0.10 | $0.20 |
430
+ | `nano-gpt/poolside/laguna-xs-2.1` | 262K | | | | | | $0.06 | $0.13 |
429
431
  | `nano-gpt/qvq-max` | 128K | | | | | | $1 | $5 |
430
432
  | `nano-gpt/qwen-3.6-plus` | 992K | | | | | | $0.33 | $2 |
431
433
  | `nano-gpt/qwen-long` | 10.0M | | | | | | $0.10 | $0.41 |
@@ -529,8 +531,8 @@ for await (const chunk of stream) {
529
531
  | `nano-gpt/TEE/gemma4-31b:thinking` | 262K | | | | | | $0.40 | $1 |
530
532
  | `nano-gpt/TEE/glm-5.1` | 203K | | | | | | $2 | $5 |
531
533
  | `nano-gpt/TEE/glm-5.1-thinking` | 203K | | | | | | $2 | $5 |
532
- | `nano-gpt/TEE/glm-5.2` | 1.0M | | | | | | $1 | $5 |
533
- | `nano-gpt/TEE/glm-5.2:thinking` | 1.0M | | | | | | $1 | $5 |
534
+ | `nano-gpt/TEE/glm-5.2` | 1.0M | | | | | | $1 | $4 |
535
+ | `nano-gpt/TEE/glm-5.2:thinking` | 1.0M | | | | | | $1 | $4 |
534
536
  | `nano-gpt/TEE/glm-5.3` | 1.0M | | | | | | $1 | $4 |
535
537
  | `nano-gpt/TEE/glm-5.3-flash` | 1.0M | | | | | | $0.15 | $0.50 |
536
538
  | `nano-gpt/TEE/gpt-oss-120b` | 131K | | | | | | $2 | $2 |
@@ -541,7 +543,6 @@ for await (const chunk of stream) {
541
543
  | `nano-gpt/TEE/llama3-3-70b` | 128K | | | | | | $2 | $3 |
542
544
  | `nano-gpt/TEE/muse-glimmer-30b` | 131K | | | | | | $0.35 | $2 |
543
545
  | `nano-gpt/TEE/qwen2.5-vl-72b-instruct` | 66K | | | | | | $0.70 | $0.70 |
544
- | `nano-gpt/TEE/qwen3-8b` | 41K | | | | | | $0.11 | $0.45 |
545
546
  | `nano-gpt/TEE/qwen3.5-27b` | 262K | | | | | | $0.30 | $2 |
546
547
  | `nano-gpt/TEE/qwen3.5-397b-a17b` | 262K | | | | | | $0.55 | $4 |
547
548
  | `nano-gpt/TEE/qwen3.6-27b` | 262K | | | | | | $0.32 | $3 |
@@ -552,6 +553,7 @@ for await (const chunk of stream) {
552
553
  | `nano-gpt/tencent/hy4-preview` | 1.0M | | | | | | $0.83 | $3 |
553
554
  | `nano-gpt/TheDrummer/Anubis-70B-v1` | 66K | | | | | | $0.31 | $0.31 |
554
555
  | `nano-gpt/TheDrummer/Anubis-70B-v1.1` | 32K | | | | | | $0.31 | $0.31 |
556
+ | `nano-gpt/TheDrummer/Artemis-v1.1` | 262K | | | | | | $0.10 | $0.45 |
555
557
  | `nano-gpt/TheDrummer/Cydonia-24B-v2` | 33K | | | | | | $0.10 | $0.12 |
556
558
  | `nano-gpt/TheDrummer/Cydonia-24B-v4` | 33K | | | | | | $0.20 | $0.24 |
557
559
  | `nano-gpt/TheDrummer/Cydonia-24B-v4.1` | 131K | | | | | | $0.35 | $0.55 |
@@ -624,7 +626,7 @@ for await (const chunk of stream) {
624
626
  | `nano-gpt/z-ai/glm-5.2:thinking` | 1.0M | | | | | | $0.42 | $1 |
625
627
  | `nano-gpt/z-ai/glm-5.3` | 1.0M | | | | | | $1 | $3 |
626
628
  | `nano-gpt/z-ai/glm-5.3-flash` | 1.0M | | | | | | $0.07 | $0.25 |
627
- | `nano-gpt/z-ai/glm-5.3-flash-uncensored` | 1.0M | | | | | | $0.35 | $1 |
629
+ | `nano-gpt/z-ai/glm-5.3-flash-uncensored` | 1.0M | | | | | | $0.07 | $0.21 |
628
630
  | `nano-gpt/z-ai/glm-5.3:thinking` | 1.0M | | | | | | $1 | $3 |
629
631
  | `nano-gpt/z-ai/glm-5v-turbo` | 203K | | | | | | $1 | $4 |
630
632
  | `nano-gpt/z-ai/glm-5v-turbo:thinking` | 203K | | | | | | $1 | $4 |
@@ -4,7 +4,7 @@
4
4
 
5
5
  # ![Ofox logo](https://models.dev/logos/ofox.svg)Ofox
6
6
 
7
- Access 114 Ofox models through Mastra's model router. Authentication is handled automatically using the `OFOX_API_KEY` environment variable.
7
+ Access 115 Ofox models through Mastra's model router. Authentication is handled automatically using the `OFOX_API_KEY` environment variable.
8
8
 
9
9
  Learn more in the [Ofox documentation](https://ofox.ai/docs).
10
10
 
@@ -125,6 +125,7 @@ for await (const chunk of stream) {
125
125
  | `ofox/openai/gpt-5.6-luna` | 1.1M | | | | | | $0.20 | $1 |
126
126
  | `ofox/openai/gpt-5.6-sol` | 1.1M | | | | | | $3 | $15 |
127
127
  | `ofox/openai/gpt-5.6-terra` | 1.1M | | | | | | $2 | $12 |
128
+ | `ofox/openai/gpt-6-astra` | 1.1M | | | | | | $10 | $50 |
128
129
  | `ofox/volcengine/doubao-seed-1-6` | 256K | | | | | | $0.12 | $0.29 |
129
130
  | `ofox/volcengine/doubao-seed-1-6-flash` | 256K | | | | | | $0.03 | $0.22 |
130
131
  | `ofox/volcengine/doubao-seed-1-6-vision` | 256K | | | | | | $0.12 | $1 |
@@ -4,7 +4,7 @@
4
4
 
5
5
  # ![SenseNova (China) logo](https://models.dev/logos/sensenova.svg)SenseNova (China)
6
6
 
7
- Access 3 SenseNova (China) models through Mastra's model router. Authentication is handled automatically using the `SENSENOVA_API_KEY` environment variable.
7
+ Access 5 SenseNova (China) models through Mastra's model router. Authentication is handled automatically using the `SENSENOVA_API_KEY` environment variable.
8
8
 
9
9
  Learn more in the [SenseNova (China) documentation](https://platform.sensenova.cn/docs).
10
10
 
@@ -39,7 +39,9 @@ for await (const chunk of stream) {
39
39
  | Model | Context | Tools | Reasoning | Image | Audio | Video | Input $/1M | Output $/1M |
40
40
  | ------------------------------------ | ------- | ----- | --------- | ----- | ----- | ----- | ---------- | ----------- |
41
41
  | `sensenova/deepseek-v4-flash` | 1.0M | | | | | | — | — |
42
+ | `sensenova/deepseek-v4-pro` | 1.0M | | | | | | — | — |
42
43
  | `sensenova/glm-5.2` | 1.0M | | | | | | — | — |
44
+ | `sensenova/kimi-k3` | 1.0M | | | | | | — | — |
43
45
  | `sensenova/sensenova-6.8-flash-lite` | 262K | | | | | | — | — |
44
46
 
45
47
  Model availability, capabilities, context windows, and pricing are sourced from [models.dev](https://models.dev) and may change.
@@ -4,7 +4,7 @@
4
4
 
5
5
  # ![Vivgrid logo](https://models.dev/logos/vivgrid.svg)Vivgrid
6
6
 
7
- Access 22 Vivgrid models through Mastra's model router. Authentication is handled automatically using the `VIVGRID_API_KEY` environment variable.
7
+ Access 27 Vivgrid models through Mastra's model router. Authentication is handled automatically using the `VIVGRID_API_KEY` environment variable.
8
8
 
9
9
  Learn more in the [Vivgrid documentation](https://docs.vivgrid.com/models).
10
10
 
@@ -19,7 +19,7 @@ const agent = new Agent({
19
19
  id: "my-agent",
20
20
  name: "My Agent",
21
21
  instructions: "You are a helpful assistant",
22
- model: "vivgrid/deepseek-v3.2"
22
+ model: "vivgrid/claude-fable-5"
23
23
  });
24
24
 
25
25
  // Generate a response
@@ -38,12 +38,16 @@ for await (const chunk of stream) {
38
38
 
39
39
  | Model | Context | Tools | Reasoning | Image | Audio | Video | Input $/1M | Output $/1M |
40
40
  | --------------------------------------- | ------- | ----- | --------- | ----- | ----- | ----- | ---------- | ----------- |
41
+ | `vivgrid/claude-fable-5` | 1.0M | | | | | | $10 | $50 |
42
+ | `vivgrid/claude-fable-5-1` | 1.0M | | | | | | $10 | $50 |
41
43
  | `vivgrid/deepseek-v3.2` | 128K | | | | | | $0.28 | $0.42 |
42
44
  | `vivgrid/deepseek-v4-flash` | 1.0M | | | | | | $0.15 | $0.30 |
43
45
  | `vivgrid/deepseek-v4-pro` | 1.0M | | | | | | $0.43 | $0.87 |
46
+ | `vivgrid/deepseek-v4-pro-0813` | 1.0M | | | | | | $1 | $3 |
44
47
  | `vivgrid/gemini-3.1-flash-lite-preview` | 1.0M | | | | | | $0.25 | $2 |
45
48
  | `vivgrid/gemini-3.1-pro-preview` | 1.0M | | | | | | $2 | $12 |
46
49
  | `vivgrid/gemini-3.7-flash` | 1.0M | | | | | | $0.75 | $4 |
50
+ | `vivgrid/gemini-3.8-flash` | 1.0M | | | | | | $0.75 | $4 |
47
51
  | `vivgrid/glm-5.2` | 1.0M | | | | | | $1 | $4 |
48
52
  | `vivgrid/glm-5.3` | 1.0M | | | | | | $1 | $4 |
49
53
  | `vivgrid/glm-5.3-flash` | 1.0M | | | | | | $0.15 | $0.50 |
@@ -59,6 +63,7 @@ for await (const chunk of stream) {
59
63
  | `vivgrid/gpt-5.6-luna` | 1.1M | | | | | | $1 | $6 |
60
64
  | `vivgrid/gpt-5.6-sol` | 1.1M | | | | | | $5 | $30 |
61
65
  | `vivgrid/gpt-5.6-terra` | 1.1M | | | | | | $3 | $15 |
66
+ | `vivgrid/gpt-6-astra` | 1.1M | | | | | | $10 | $50 |
62
67
  | `vivgrid/kimi-k3` | 1.0M | | | | | | $3 | $15 |
63
68
 
64
69
  Model availability, capabilities, context windows, and pricing are sourced from [models.dev](https://models.dev) and may change.
@@ -73,7 +78,7 @@ const agent = new Agent({
73
78
  name: "custom-agent",
74
79
  model: {
75
80
  url: "https://api.vivgrid.com/v1",
76
- id: "vivgrid/deepseek-v3.2",
81
+ id: "vivgrid/claude-fable-5",
77
82
  apiKey: process.env.VIVGRID_API_KEY,
78
83
  headers: {
79
84
  "X-Custom-Header": "value"
@@ -92,7 +97,7 @@ const agent = new Agent({
92
97
  const useAdvanced = requestContext.task === "complex";
93
98
  return useAdvanced
94
99
  ? "vivgrid/kimi-k3"
95
- : "vivgrid/deepseek-v3.2";
100
+ : "vivgrid/claude-fable-5";
96
101
  }
97
102
  });
98
103
  ```
@@ -43,7 +43,7 @@ cleanup()
43
43
 
44
44
  ### Using the `durable` config flag
45
45
 
46
- Set `durable: true` on `AgentConfig` and the agent is automatically wrapped with `createDurableAgent` when it's attached to a `Mastra` instance. Use an object to forward advanced options such as `cache`, `pubsub`, `maxSteps`, or `cleanupTimeoutMs`.
46
+ Set `durable: true` on `AgentConfig` and the agent is automatically wrapped with `createDurableAgent` when it's attached to a `Mastra` instance. Use an object to forward advanced options such as `cache`, `pubsub`, `maxSteps`, `cleanupTimeoutMs`, or `shouldCache`.
47
47
 
48
48
  ```typescript
49
49
  import { Mastra } from '@mastra/core'
@@ -90,6 +90,8 @@ Returns: `DurableAgent`
90
90
 
91
91
  **maxSteps** (`number`): Maximum number of steps for the agentic loop.
92
92
 
93
+ **shouldCache** (`(topic: string) => boolean`): Per-topic opt-out of the replay cache. Return false to publish a topic straight to the underlying PubSub without recording it; subscribers of that topic receive live events only and cannot resume from an offset. Useful for trading replay for minimum publish latency on hot topics when the cache is remote (for example, cross-region Redis). Run-local topics are always excluded, regardless of this option.
94
+
93
95
  ## `createEventedAgent(options)`
94
96
 
95
97
  Wraps an `Agent` with fire-and-forget durable execution on the built-in workflow engine. Like `createDurableAgent`, it returns a result you stream from, but the underlying workflow runs non-blocking (via `startAsync`) instead of running to completion before the stream is wired up. Use it when you want the run to progress independently of the caller. It doesn't accept `id` or `name` overrides.
@@ -112,6 +114,8 @@ Returns: `EventedAgent` (a subclass of `DurableAgent`)
112
114
 
113
115
  **maxSteps** (`number`): Maximum number of steps for the agentic loop.
114
116
 
117
+ **shouldCache** (`(topic: string) => boolean`): Per-topic opt-out of the replay cache. Return false to publish a topic straight to the underlying PubSub without recording it; subscribers of that topic receive live events only and cannot resume from an offset. Useful for trading replay for minimum publish latency on hot topics when the cache is remote (for example, cross-region Redis). Run-local topics are always excluded, regardless of this option.
118
+
115
119
  ## Constructor parameters
116
120
 
117
121
  The `DurableAgent` class accepts the same options as `createDurableAgent`, plus `cleanupTimeoutMs`. Prefer the factory unless you need to subclass.
@@ -128,6 +132,8 @@ The `DurableAgent` class accepts the same options as `createDurableAgent`, plus
128
132
 
129
133
  **maxSteps** (`number`): Maximum number of steps for the agentic loop.
130
134
 
135
+ **shouldCache** (`(topic: string) => boolean`): Per-topic opt-out of the replay cache. Return false to publish a topic straight to the underlying PubSub without recording it; subscribers of that topic receive live events only and cannot resume from an offset. Useful for trading replay for minimum publish latency on hot topics when the cache is remote (for example, cross-region Redis). Run-local topics are always excluded, regardless of this option.
136
+
131
137
  **cleanupTimeoutMs** (`number`): Grace period in milliseconds before registry entries are cleaned up automatically after a stream finishes or errors. Set to 0 to disable auto-cleanup and require a manual cleanup() call. Auto-cleanup does not fire on suspended events. (Default: `30000`)
132
138
 
133
139
  ## Methods
@@ -23,7 +23,7 @@ Saving changed snapshot fields creates a new latest version. Saving identical sn
23
23
 
24
24
  If an active version exists, creating a draft doesn't change the version handling published requests. Publishing updates `activeVersionId`. Restoring a historical version copies its configuration into a new inactive draft.
25
25
 
26
- The direct namespace methods and REST APIs differ in one important way. `editor.prompt.update()` creates an inactive draft. `editor.agent.update()` creates a version and immediately assigns it to `activeVersionId`. The stored-agent REST `PATCH` route creates an inactive draft unless `autoPublish` is enabled.
26
+ The direct namespace methods and the REST APIs agree on this. `editor.prompt.update()` and `editor.agent.update()` both create an inactive draft. Neither assigns the new version to `activeVersionId`. To publish a version from the SDK, pass it explicitly: `editor.agent.update({ id, status: 'published', activeVersionId: version.id })`. The stored-agent REST `PATCH` route also creates an inactive draft unless `autoPublish` is enabled, and `POST /stored/agents/:id/versions/:versionId/activate` publishes it.
27
27
 
28
28
  When a generic stored resource has no active version, published resolution can fall back to the latest snapshot. For a code-defined agent override, requesting `status: 'published'` without an active override returns the original code agent.
29
29
 
@@ -61,7 +61,7 @@ const graphTool = createGraphRAGTool({
61
61
 
62
62
  The tool returns an object with:
63
63
 
64
- **relevantContext** (`string`): Combined text from the most relevant document chunks, retrieved using graph-based ranking
64
+ **relevantContext** (`string[]`): Array of chunk text strings for the most relevant document chunks, in rank order, retrieved using graph-based ranking. Text is read from the text field of each chunk's metadata.
65
65
 
66
66
  **sources** (`QueryResult[]`): Array of full retrieval result objects. Each object contains all information needed to reference the original document, chunk, and similarity score.
67
67
 
@@ -77,6 +77,8 @@ The tool returns an object with:
77
77
  }
78
78
  ```
79
79
 
80
+ The graph is built from the `text` field of each result's metadata, so `document` (and `relevantContext`) contain that text. Store the chunk text under `metadata.text` when upserting.
81
+
80
82
  ## Default tool description
81
83
 
82
84
  The default description focuses on:
@@ -79,7 +79,7 @@ const queryTool = createVectorQueryTool({
79
79
 
80
80
  The tool returns an object with:
81
81
 
82
- **relevantContext** (`string`): Combined text from the most relevant document chunks
82
+ **relevantContext** (`any[]`): Array of metadata objects for the most relevant chunks, in rank order (one entry per result). The chunk text is available at relevantContext\[i].text when it was stored in metadata during ingestion.
83
83
 
84
84
  **sources** (`QueryResult[]`): Array of full retrieval result objects. Each object contains all information needed to reference the original document, chunk, and similarity score.
85
85
 
@@ -95,6 +95,8 @@ The tool returns an object with:
95
95
  }
96
96
  ```
97
97
 
98
+ `document` is only populated by vector stores whose `query()` returns document content, such as Chroma, Elasticsearch, LanceDB, and MongoDB. For other stores, such as PgVector, it's an empty string. Read the chunk text from `sources[i].metadata.text` (or whichever metadata key you stored it under).
99
+
98
100
  ## Default tool description
99
101
 
100
102
  The default description focuses on:
@@ -489,7 +491,7 @@ The tool is created with:
489
491
 
490
492
  - **ID**: `VectorQuery {vectorStoreName} {indexName} Tool`
491
493
  - **Input Schema**: Requires queryText and filter objects
492
- - **Output Schema**: Returns relevantContext string
494
+ - **Output Schema**: Returns `relevantContext` (array of chunk metadata) and `sources` (array of `QueryResult`)
493
495
 
494
496
  ## Related
495
497
 
package/package.json CHANGED
@@ -1,6 +1,6 @@
1
1
  {
2
2
  "name": "@mastra/mcp-docs-server",
3
- "version": "1.2.24-alpha.14",
3
+ "version": "1.2.24-alpha.18",
4
4
  "description": "MCP server for accessing Mastra.ai documentation, changelogs, and news.",
5
5
  "type": "module",
6
6
  "main": "dist/index.js",
@@ -27,8 +27,8 @@
27
27
  "jsdom": "^26.1.0",
28
28
  "local-pkg": "^1.1.2",
29
29
  "zod": "^4.4.3",
30
- "@mastra/mcp": "^1.17.3",
31
- "@mastra/core": "1.65.0-alpha.7"
30
+ "@mastra/core": "1.65.0-alpha.9",
31
+ "@mastra/mcp": "^1.17.3"
32
32
  },
33
33
  "devDependencies": {
34
34
  "@hono/node-server": "^2.0.0",
@@ -44,9 +44,9 @@
44
44
  "tsx": "^4.23.1",
45
45
  "typescript": "^7.0.2",
46
46
  "vitest": "4.1.10",
47
- "@internal/types-builder": "0.0.105",
48
47
  "@internal/lint": "0.0.130",
49
- "@mastra/core": "1.65.0-alpha.7"
48
+ "@internal/types-builder": "0.0.105",
49
+ "@mastra/core": "1.65.0-alpha.9"
50
50
  },
51
51
  "homepage": "https://mastra.ai",
52
52
  "repository": {