@mastra/mcp-docs-server 1.2.24-alpha.14 → 1.2.24-alpha.18
This diff represents the content of publicly available package versions that have been released to one of the supported registries. The information contained in this diff is provided for informational purposes only and reflects changes between package versions as they appear in their respective public registries.
- package/.docs/docs/harness/durable-agents.md +11 -0
- package/.docs/integrations/observability/langfuse.md +3 -0
- package/.docs/models/gateways/neon.md +5 -1
- package/.docs/models/gateways/netlify.md +1 -2
- package/.docs/models/gateways/openrouter.md +1 -4
- package/.docs/models/gateways/vercel.md +1 -3
- package/.docs/models/index.md +1 -1
- package/.docs/models/providers/cerebras.md +5 -5
- package/.docs/models/providers/cortecs.md +1 -5
- package/.docs/models/providers/crof.md +27 -26
- package/.docs/models/providers/deepinfra.md +1 -1
- package/.docs/models/providers/edenai.md +5 -4
- package/.docs/models/providers/fireworks-ai.md +2 -1
- package/.docs/models/providers/hyper.md +3 -3
- package/.docs/models/providers/kilo.md +8 -10
- package/.docs/models/providers/llmgateway-providers.md +1 -2
- package/.docs/models/providers/moonshotai-cn.md +4 -10
- package/.docs/models/providers/moonshotai.md +4 -10
- package/.docs/models/providers/nano-gpt.md +10 -8
- package/.docs/models/providers/ofox.md +2 -1
- package/.docs/models/providers/sensenova.md +3 -1
- package/.docs/models/providers/vivgrid.md +9 -4
- package/.docs/reference/agents/durable-agent.md +7 -1
- package/.docs/reference/editor/versioning.md +1 -1
- package/.docs/reference/tools/graph-rag-tool.md +3 -1
- package/.docs/reference/tools/vector-query-tool.md +4 -2
- package/package.json +5 -5
|
@@ -189,6 +189,17 @@ export const durableAgent = createDurableAgent({
|
|
|
189
189
|
|
|
190
190
|
`createInngestAgent()` doesn't enable caching by default. Pass a `cache` option or register the agent with a `Mastra` instance that has a `serverCache` configured to enable resumable streams.
|
|
191
191
|
|
|
192
|
+
For cached topics, each published event is recorded in the cache before it's delivered live, so publish latency is bounded by the round-trip to the cache. If the cache write fails, the event is still delivered live but can't be replayed. `@mastra/redis` and `@mastra/valkey` record each event in a single round-trip using a Lua script. `@mastra/redis` falls back to separate commands when the client has no `evalScript` or when Redis Cluster rejects the multi-key script. Per-run workflow watch events (`workflow.events.v2.*`) are never cached. If the cache is remote (for example, in another region) and you don't need to resume a topic, use `shouldCache` to publish that topic straight through:
|
|
193
|
+
|
|
194
|
+
```typescript
|
|
195
|
+
export const durableAgent = createDurableAgent({
|
|
196
|
+
agent,
|
|
197
|
+
cache,
|
|
198
|
+
// Skip the replay cache for the per-chunk stream topic; other topics stay resumable.
|
|
199
|
+
shouldCache: topic => !topic.startsWith('agent.stream.'),
|
|
200
|
+
})
|
|
201
|
+
```
|
|
202
|
+
|
|
192
203
|
## Streaming with background tasks
|
|
193
204
|
|
|
194
205
|
Durable agents support the same [`untilIdle`](https://mastra.ai/reference/streaming/agents/stream) option as regular agents. When `untilIdle` is set, `stream()` keeps the connection open across background-task continuations until the agent is idle:
|
|
@@ -216,10 +216,13 @@ const tracingOptions = {
|
|
|
216
216
|
|
|
217
217
|
This example produces `langfuse.trace.metadata.customerId` and `langfuse.trace.metadata.tier`.
|
|
218
218
|
|
|
219
|
+
Metadata on the root span is also forwarded. Mastra sets `runId` and `resourceId` on every agent and workflow root span, and you can add your own keys through `tracingOptions.metadata`. The exporter forwards each of these root span keys to `langfuse.trace.metadata.<key>`. Keys that map to a dedicated Langfuse field (`userId`, `sessionId`, `threadId`, `traceName`, and `version`) are not duplicated as trace metadata. Metadata on child spans stays on the observation and never changes the trace.
|
|
220
|
+
|
|
219
221
|
Notes:
|
|
220
222
|
|
|
221
223
|
- The reserved `prompt` key is used for [prompt linking](#prompt-linking) and isn't forwarded as trace metadata.
|
|
222
224
|
- The reserved identity keys `agentId`, `agentName`, `workflowId`, and `workflowName` are set from the root span and take precedence over custom values with the same name.
|
|
225
|
+
- Keys under `langfuse` take precedence over root span metadata with the same name.
|
|
223
226
|
- Values are sent as strings, because Langfuse maps trace metadata attributes as strings. Numbers, booleans, and objects are serialized with JSON. Langfuse Cloud restores them to their original types on ingestion.
|
|
224
227
|
|
|
225
228
|
## Prompt linking
|
|
@@ -4,7 +4,7 @@
|
|
|
4
4
|
|
|
5
5
|
# Neon
|
|
6
6
|
|
|
7
|
-
Neon aggregates models from multiple providers with enhanced features like rate limiting and failover. Access
|
|
7
|
+
Neon aggregates models from multiple providers with enhanced features like rate limiting and failover. Access 46 models through Mastra's model router.
|
|
8
8
|
|
|
9
9
|
Learn more in the [Neon documentation](https://neon.com/docs).
|
|
10
10
|
|
|
@@ -36,6 +36,7 @@ NEON_AI_GATEWAY_TOKEN=your-gateway-key
|
|
|
36
36
|
| Model |
|
|
37
37
|
| ----------------------------- |
|
|
38
38
|
| `claude-fable-5` |
|
|
39
|
+
| `claude-fable-5-1` |
|
|
39
40
|
| `claude-haiku-4-5` |
|
|
40
41
|
| `claude-opus-4-1` |
|
|
41
42
|
| `claude-opus-4-5` |
|
|
@@ -54,6 +55,7 @@ NEON_AI_GATEWAY_TOKEN=your-gateway-key
|
|
|
54
55
|
| `gemini-3-flash` |
|
|
55
56
|
| `gemma-3-12b` |
|
|
56
57
|
| `glm-5-2` |
|
|
58
|
+
| `glm-5-3-flash` |
|
|
57
59
|
| `gpt-5` |
|
|
58
60
|
| `gpt-5-1` |
|
|
59
61
|
| `gpt-5-2` |
|
|
@@ -68,8 +70,10 @@ NEON_AI_GATEWAY_TOKEN=your-gateway-key
|
|
|
68
70
|
| `gpt-5-6-terra` |
|
|
69
71
|
| `gpt-5-mini` |
|
|
70
72
|
| `gpt-5-nano` |
|
|
73
|
+
| `gpt-6-astra` |
|
|
71
74
|
| `gpt-oss-120b` |
|
|
72
75
|
| `gpt-oss-20b` |
|
|
76
|
+
| `grok-4-6` |
|
|
73
77
|
| `inkling` |
|
|
74
78
|
| `kimi-k3` |
|
|
75
79
|
| `llama-4-maverick` |
|
|
@@ -4,7 +4,7 @@
|
|
|
4
4
|
|
|
5
5
|
# Netlify
|
|
6
6
|
|
|
7
|
-
Netlify AI Gateway provides unified access to multiple providers with built-in caching and observability. Access
|
|
7
|
+
Netlify AI Gateway provides unified access to multiple providers with built-in caching and observability. Access 242 models through Mastra's model router.
|
|
8
8
|
|
|
9
9
|
Learn more in the [Netlify documentation](https://docs.netlify.com/build/ai-gateway/overview/).
|
|
10
10
|
|
|
@@ -279,6 +279,5 @@ ANTHROPIC_API_KEY=ant-...
|
|
|
279
279
|
| `openrouter/z-ai/glm-5` |
|
|
280
280
|
| `openrouter/z-ai/glm-5.1` |
|
|
281
281
|
| `openrouter/z-ai/glm-5.2` |
|
|
282
|
-
| `openrouter/z-ai/glm-5.2:free` |
|
|
283
282
|
| `openrouter/z-ai/glm-5.3` |
|
|
284
283
|
| `openrouter/z-ai/glm-5.3-flash` |
|
|
@@ -4,7 +4,7 @@
|
|
|
4
4
|
|
|
5
5
|
# OpenRouter
|
|
6
6
|
|
|
7
|
-
OpenRouter aggregates models from multiple providers with enhanced features like rate limiting and failover. Access
|
|
7
|
+
OpenRouter aggregates models from multiple providers with enhanced features like rate limiting and failover. Access 358 models through Mastra's model router.
|
|
8
8
|
|
|
9
9
|
Learn more in the [OpenRouter documentation](https://openrouter.ai/models).
|
|
10
10
|
|
|
@@ -174,9 +174,7 @@ ANTHROPIC_API_KEY=ant-...
|
|
|
174
174
|
| `minimax/minimax-m2.1` |
|
|
175
175
|
| `minimax/minimax-m2.5` |
|
|
176
176
|
| `minimax/minimax-m2.7` |
|
|
177
|
-
| `minimax/minimax-m2.7:free` |
|
|
178
177
|
| `minimax/minimax-m3` |
|
|
179
|
-
| `minimax/minimax-m3:free` |
|
|
180
178
|
| `mistralai/codestral-2508` |
|
|
181
179
|
| `mistralai/devstral-2512` |
|
|
182
180
|
| `mistralai/ministral-14b-2512` |
|
|
@@ -395,7 +393,6 @@ ANTHROPIC_API_KEY=ant-...
|
|
|
395
393
|
| `z-ai/glm-5-turbo` |
|
|
396
394
|
| `z-ai/glm-5.1` |
|
|
397
395
|
| `z-ai/glm-5.2` |
|
|
398
|
-
| `z-ai/glm-5.2:free` |
|
|
399
396
|
| `z-ai/glm-5.3` |
|
|
400
397
|
| `z-ai/glm-5.3-flash` |
|
|
401
398
|
| `z-ai/glm-5v-turbo` |
|
|
@@ -4,7 +4,7 @@
|
|
|
4
4
|
|
|
5
5
|
# Vercel
|
|
6
6
|
|
|
7
|
-
Vercel aggregates models from multiple providers with enhanced features like rate limiting and failover. Access
|
|
7
|
+
Vercel aggregates models from multiple providers with enhanced features like rate limiting and failover. Access 371 models through Mastra's model router.
|
|
8
8
|
|
|
9
9
|
Learn more in the [Vercel documentation](https://ai-sdk.dev/providers/ai-sdk-providers).
|
|
10
10
|
|
|
@@ -221,10 +221,8 @@ ANTHROPIC_API_KEY=ant-...
|
|
|
221
221
|
| `minimax/minimax-m2.5` |
|
|
222
222
|
| `minimax/minimax-m2.5-highspeed` |
|
|
223
223
|
| `minimax/minimax-m2.7` |
|
|
224
|
-
| `minimax/minimax-m2.7-free` |
|
|
225
224
|
| `minimax/minimax-m2.7-highspeed` |
|
|
226
225
|
| `minimax/minimax-m3` |
|
|
227
|
-
| `minimax/minimax-m3-free` |
|
|
228
226
|
| `mistral/codestral` |
|
|
229
227
|
| `mistral/codestral-embed` |
|
|
230
228
|
| `mistral/devstral-2` |
|
package/.docs/models/index.md
CHANGED
|
@@ -4,7 +4,7 @@
|
|
|
4
4
|
|
|
5
5
|
# Model Providers
|
|
6
6
|
|
|
7
|
-
Mastra provides a unified interface for working with LLMs across multiple providers, giving you access to
|
|
7
|
+
Mastra provides a unified interface for working with LLMs across multiple providers, giving you access to 7098 models from 200 providers through a single API.
|
|
8
8
|
|
|
9
9
|
## Features
|
|
10
10
|
|
|
@@ -19,7 +19,7 @@ const agent = new Agent({
|
|
|
19
19
|
id: "my-agent",
|
|
20
20
|
name: "My Agent",
|
|
21
21
|
instructions: "You are a helpful assistant",
|
|
22
|
-
model: "cerebras/
|
|
22
|
+
model: "cerebras/gpt-oss-120b"
|
|
23
23
|
});
|
|
24
24
|
|
|
25
25
|
// Generate a response
|
|
@@ -36,8 +36,8 @@ for await (const chunk of stream) {
|
|
|
36
36
|
|
|
37
37
|
| Model | Context | Tools | Reasoning | Image | Audio | Video | Input $/1M | Output $/1M |
|
|
38
38
|
| ----------------------- | ------- | ----- | --------- | ----- | ----- | ----- | ---------- | ----------- |
|
|
39
|
-
| `cerebras/gemma-4-31b` | 131K | | | | | | $0.99 | $1 |
|
|
40
39
|
| `cerebras/gpt-oss-120b` | 131K | | | | | | $0.35 | $0.75 |
|
|
40
|
+
| `cerebras/qwen-3.8-27b` | 66K | | | | | | $0.99 | $1 |
|
|
41
41
|
|
|
42
42
|
Model availability, capabilities, context windows, and pricing are sourced from [models.dev](https://models.dev) and may change.
|
|
43
43
|
|
|
@@ -50,7 +50,7 @@ const agent = new Agent({
|
|
|
50
50
|
id: "custom-agent",
|
|
51
51
|
name: "custom-agent",
|
|
52
52
|
model: {
|
|
53
|
-
id: "cerebras/
|
|
53
|
+
id: "cerebras/gpt-oss-120b",
|
|
54
54
|
apiKey: process.env.CEREBRAS_API_KEY,
|
|
55
55
|
headers: {
|
|
56
56
|
"X-Custom-Header": "value"
|
|
@@ -68,8 +68,8 @@ const agent = new Agent({
|
|
|
68
68
|
model: ({ requestContext }) => {
|
|
69
69
|
const useAdvanced = requestContext.task === "complex";
|
|
70
70
|
return useAdvanced
|
|
71
|
-
? "cerebras/
|
|
72
|
-
: "cerebras/
|
|
71
|
+
? "cerebras/qwen-3.8-27b"
|
|
72
|
+
: "cerebras/gpt-oss-120b";
|
|
73
73
|
}
|
|
74
74
|
});
|
|
75
75
|
```
|
|
@@ -4,7 +4,7 @@
|
|
|
4
4
|
|
|
5
5
|
# Cortecs
|
|
6
6
|
|
|
7
|
-
Access
|
|
7
|
+
Access 105 Cortecs models through Mastra's model router. Authentication is handled automatically using the `CORTECS_API_KEY` environment variable.
|
|
8
8
|
|
|
9
9
|
Learn more in the [Cortecs documentation](https://cortecs.ai).
|
|
10
10
|
|
|
@@ -50,7 +50,6 @@ for await (const chunk of stream) {
|
|
|
50
50
|
| `cortecs/claude-sonnet-4` | 200K | | | | | | $3 | $14 |
|
|
51
51
|
| `cortecs/claude-sonnet-5` | 1.0M | | | | | | $2 | $11 |
|
|
52
52
|
| `cortecs/codestral-2508` | 256K | | | | | | $0.33 | $1 |
|
|
53
|
-
| `cortecs/cosmos3-super-reasoner` | 256K | | | | | | $0.10 | $0.30 |
|
|
54
53
|
| `cortecs/deepseek-r1-0528` | 164K | | | | | | $0.65 | $3 |
|
|
55
54
|
| `cortecs/deepseek-v3.2` | 164K | | | | | | $0.30 | $0.49 |
|
|
56
55
|
| `cortecs/deepseek-v4-flash-0731` | 1.0M | | | | | | $0.13 | $0.28 |
|
|
@@ -98,7 +97,6 @@ for await (const chunk of stream) {
|
|
|
98
97
|
| `cortecs/kimi-k3` | 1.0M | | | | | | $3 | $15 |
|
|
99
98
|
| `cortecs/llama-3.1-405b-instruct` | 128K | | | | | | $2 | $2 |
|
|
100
99
|
| `cortecs/llama-3.1-8b-instruct` | 128K | | | | | | $0.17 | $0.17 |
|
|
101
|
-
| `cortecs/llama-3.1-nemotron-ultra-253b-v1` | 128K | | | | | | $0.60 | $2 |
|
|
102
100
|
| `cortecs/llama-3.3-70b-instruct` | 131K | | | | | | $0.13 | $0.40 |
|
|
103
101
|
| `cortecs/minicpm-v-4.5` | 32K | | | | | | $0.65 | $1 |
|
|
104
102
|
| `cortecs/minimax-m2` | 400K | | | | | | $0.35 | $1 |
|
|
@@ -125,7 +123,6 @@ for await (const chunk of stream) {
|
|
|
125
123
|
| `cortecs/nova-micro-v1` | 128K | | | | | | $0.04 | $0.16 |
|
|
126
124
|
| `cortecs/nova-pro-v1` | 300K | | | | | | $0.92 | $4 |
|
|
127
125
|
| `cortecs/nvidia-nemotron-3-nano-30b-a3b` | 256K | | | | | | $0.06 | $0.24 |
|
|
128
|
-
| `cortecs/nvidia-nemotron-3-nano-omni` | 300K | | | | | | $0.06 | $0.24 |
|
|
129
126
|
| `cortecs/pixtral-12b-2409` | 128K | | | | | | $0.22 | $0.22 |
|
|
130
127
|
| `cortecs/pixtral-large-2502` | 128K | | | | | | $2 | $6 |
|
|
131
128
|
| `cortecs/qwen2.5-vl-72b-instruct` | 32K | | | | | | $1 | $1 |
|
|
@@ -134,7 +131,6 @@ for await (const chunk of stream) {
|
|
|
134
131
|
| `cortecs/qwen3-32b` | 32K | | | | | | $0.09 | $0.31 |
|
|
135
132
|
| `cortecs/qwen3-coder-30b-a3b-instruct` | 262K | | | | | | $0.07 | $0.24 |
|
|
136
133
|
| `cortecs/qwen3-coder-next` | 256K | | | | | | $0.17 | $0.89 |
|
|
137
|
-
| `cortecs/qwen3-next-80b-a3b-thinking` | 128K | | | | | | $0.15 | $1 |
|
|
138
134
|
| `cortecs/qwen3-vl-235b-a22b` | 256K | | | | | | $0.62 | $3 |
|
|
139
135
|
| `cortecs/qwen3.5-122b-a10b` | 262K | | | | | | $0.49 | $3 |
|
|
140
136
|
| `cortecs/qwen3.5-397b-a17b` | 262K | | | | | | $0.67 | $4 |
|
|
@@ -4,7 +4,7 @@
|
|
|
4
4
|
|
|
5
5
|
# CrofAI
|
|
6
6
|
|
|
7
|
-
Access
|
|
7
|
+
Access 24 CrofAI models through Mastra's model router. Authentication is handled automatically using the `CROF_API_KEY` environment variable.
|
|
8
8
|
|
|
9
9
|
Learn more in the [CrofAI documentation](https://crof.ai/docs).
|
|
10
10
|
|
|
@@ -36,31 +36,32 @@ for await (const chunk of stream) {
|
|
|
36
36
|
|
|
37
37
|
## Models
|
|
38
38
|
|
|
39
|
-
| Model
|
|
40
|
-
|
|
|
41
|
-
| `crof/deepseek-v3.2`
|
|
42
|
-
| `crof/deepseek-v4-flash`
|
|
43
|
-
| `crof/deepseek-v4-flash-0731`
|
|
44
|
-
| `crof/deepseek-v4-
|
|
45
|
-
| `crof/deepseek-v4-pro
|
|
46
|
-
| `crof/
|
|
47
|
-
| `crof/
|
|
48
|
-
| `crof/glm-5.
|
|
49
|
-
| `crof/glm-5.
|
|
50
|
-
| `crof/glm-5.3
|
|
51
|
-
| `crof/
|
|
52
|
-
| `crof/greg-
|
|
53
|
-
| `crof/greg-2-
|
|
54
|
-
| `crof/greg-
|
|
55
|
-
| `crof/
|
|
56
|
-
| `crof/kimi-k2.
|
|
57
|
-
| `crof/kimi-
|
|
58
|
-
| `crof/kimi-k3
|
|
59
|
-
| `crof/
|
|
60
|
-
| `crof/
|
|
61
|
-
| `crof/qwen3.5-
|
|
62
|
-
| `crof/qwen3.
|
|
63
|
-
| `crof/qwen3.
|
|
39
|
+
| Model | Context | Tools | Reasoning | Image | Audio | Video | Input $/1M | Output $/1M |
|
|
40
|
+
| ----------------------------------- | ------- | ----- | --------- | ----- | ----- | ----- | ---------- | ----------- |
|
|
41
|
+
| `crof/deepseek-v3.2` | 164K | | | | | | $0.18 | $0.35 |
|
|
42
|
+
| `crof/deepseek-v4-flash` | 1.0M | | | | | | $0.12 | $0.21 |
|
|
43
|
+
| `crof/deepseek-v4-flash-0731` | 1.0M | | | | | | $0.08 | $0.10 |
|
|
44
|
+
| `crof/deepseek-v4-flash-vision-exp` | 1.0M | | | | | | $0.08 | $0.20 |
|
|
45
|
+
| `crof/deepseek-v4-pro` | 1.0M | | | | | | $0.35 | $0.80 |
|
|
46
|
+
| `crof/deepseek-v4-pro-0813` | 1.0M | | | | | | $0.35 | $0.80 |
|
|
47
|
+
| `crof/gemma-4-31b-it` | 262K | | | | | | $0.10 | $0.30 |
|
|
48
|
+
| `crof/glm-5.1` | 203K | | | | | | $0.45 | $2 |
|
|
49
|
+
| `crof/glm-5.2` | 1.0M | | | | | | $0.30 | $1 |
|
|
50
|
+
| `crof/glm-5.3` | 1.0M | | | | | | $0.40 | $1 |
|
|
51
|
+
| `crof/glm-5.3-flash` | 1.0M | | | | | | $0.07 | $0.22 |
|
|
52
|
+
| `crof/greg-1-mini` | 229K | | | | | | $0.07 | $0.15 |
|
|
53
|
+
| `crof/greg-2-super` | 229K | | | | | | $2 | $5 |
|
|
54
|
+
| `crof/greg-2-ultra` | 229K | | | | | | $3 | $10 |
|
|
55
|
+
| `crof/greg-rp` | 229K | | | | | | $0.10 | $0.30 |
|
|
56
|
+
| `crof/kimi-k2.6` | 262K | | | | | | $0.50 | $2 |
|
|
57
|
+
| `crof/kimi-k2.7-code` | 262K | | | | | | $0.55 | $2 |
|
|
58
|
+
| `crof/kimi-k3` | 1.0M | | | | | | $2 | $8 |
|
|
59
|
+
| `crof/kimi-k3-eco` | 1.0M | | | | | | $1 | $4 |
|
|
60
|
+
| `crof/mimo-v2.5-pro` | 1.0M | | | | | | $0.40 | $0.80 |
|
|
61
|
+
| `crof/qwen3.5-397b-a17b` | 262K | | | | | | $0.35 | $2 |
|
|
62
|
+
| `crof/qwen3.5-9b` | 262K | | | | | | $0.04 | $0.15 |
|
|
63
|
+
| `crof/qwen3.6-27b` | 262K | | | | | | $0.20 | $2 |
|
|
64
|
+
| `crof/qwen3.8-27b` | 262K | | | | | | $0.20 | $2 |
|
|
64
65
|
|
|
65
66
|
Model availability, capabilities, context windows, and pricing are sourced from [models.dev](https://models.dev) and may change.
|
|
66
67
|
|
|
@@ -45,7 +45,7 @@ for await (const chunk of stream) {
|
|
|
45
45
|
| `deepinfra/deepseek-ai/DeepSeek-V3.1` | 164K | | | | | | $0.25 | $0.95 |
|
|
46
46
|
| `deepinfra/deepseek-ai/DeepSeek-V3.2` | 164K | | | | | | $0.26 | $0.38 |
|
|
47
47
|
| `deepinfra/deepseek-ai/DeepSeek-V4-Flash` | 1.0M | | | | | | $0.09 | $0.18 |
|
|
48
|
-
| `deepinfra/deepseek-ai/DeepSeek-V4-Flash-0731` | 1.0M | | | | | | $0.
|
|
48
|
+
| `deepinfra/deepseek-ai/DeepSeek-V4-Flash-0731` | 1.0M | | | | | | $0.06 | $0.18 |
|
|
49
49
|
| `deepinfra/deepseek-ai/DeepSeek-V4-Flash-Vision-Exp` | 1.0M | | | | | | $0.44 | $1 |
|
|
50
50
|
| `deepinfra/deepseek-ai/DeepSeek-V4-Pro` | 1.0M | | | | | | $1 | $3 |
|
|
51
51
|
| `deepinfra/deepseek-ai/DeepSeek-V4-Pro-0813` | 1.0M | | | | | | $1 | $3 |
|
|
@@ -4,7 +4,7 @@
|
|
|
4
4
|
|
|
5
5
|
# Eden AI
|
|
6
6
|
|
|
7
|
-
Access
|
|
7
|
+
Access 250 Eden AI models through Mastra's model router. Authentication is handled automatically using the `EDENAI_API_KEY` environment variable.
|
|
8
8
|
|
|
9
9
|
Learn more in the [Eden AI documentation](https://docs.edenai.co).
|
|
10
10
|
|
|
@@ -78,12 +78,13 @@ for await (const chunk of stream) {
|
|
|
78
78
|
| `edenai/databricks/databricks-gpt-oss-120b@eu` | 131K | | | | | | $0.15 | $0.60 |
|
|
79
79
|
| `edenai/databricks/databricks-gpt-oss-20b` | 131K | | | | | | $0.07 | $0.30 |
|
|
80
80
|
| `edenai/databricks/databricks-gpt-oss-20b@eu` | 131K | | | | | | $0.07 | $0.30 |
|
|
81
|
+
| `edenai/databricks/databricks-inkling` | 1.0M | | | | | | $1 | $4 |
|
|
81
82
|
| `edenai/deepinfra/ByteDance/Seed-2.0-code` | 256K | | | | | | $0.50 | $3 |
|
|
82
83
|
| `edenai/deepinfra/ByteDance/Seed-2.0-mini` | 256K | | | | | | $0.10 | $0.40 |
|
|
83
84
|
| `edenai/deepinfra/deepseek-ai/DeepSeek-R1` | 164K | | | | | | $0.70 | $2 |
|
|
84
85
|
| `edenai/deepinfra/deepseek-ai/DeepSeek-V3` | 164K | | | | | | $0.32 | $0.89 |
|
|
85
86
|
| `edenai/deepinfra/deepseek-ai/DeepSeek-V3-0324` | 164K | | | | | | $0.24 | $0.90 |
|
|
86
|
-
| `edenai/deepinfra/deepseek-ai/DeepSeek-V4-Flash-0731` | 1.0M | | | | | | $0.
|
|
87
|
+
| `edenai/deepinfra/deepseek-ai/DeepSeek-V4-Flash-0731` | 1.0M | | | | | | $0.06 | $0.18 |
|
|
87
88
|
| `edenai/deepinfra/deepseek-ai/DeepSeek-V4-Pro-0813` | 1.0M | | | | | | $1 | $3 |
|
|
88
89
|
| `edenai/deepinfra/meta-llama/Llama-3.2-11B-Vision-Instruct` | 131K | | | | | | $0.34 | $0.34 |
|
|
89
90
|
| `edenai/deepinfra/meta-llama/Llama-3.3-70B-Instruct` | 131K | | | | | | $0.10 | $0.32 |
|
|
@@ -110,7 +111,7 @@ for await (const chunk of stream) {
|
|
|
110
111
|
| `edenai/fireworks_ai/accounts/fireworks/models/inkling` | 1.0M | | | | | | $1 | $4 |
|
|
111
112
|
| `edenai/fireworks_ai/accounts/fireworks/models/muse-glimmer-30b` | 131K | | | | | | $0.35 | $2 |
|
|
112
113
|
| `edenai/fireworks_ai/gpt-oss-120b` | 131K | | | | | | $0.15 | $0.60 |
|
|
113
|
-
| `edenai/flexai/
|
|
114
|
+
| `edenai/flexai/DeepSeek-V4-Flash-0731` | 786K | | | | | | $0.03 | $0.10 |
|
|
114
115
|
| `edenai/flexai/gpt-oss-120b` | 131K | | | | | | $0.04 | $0.10 |
|
|
115
116
|
| `edenai/flexai/gpt-oss-20b` | 131K | | | | | | $0.02 | $0.10 |
|
|
116
117
|
| `edenai/flexai/Muse-Glimmer-30B` | 131K | | | | | | $0.30 | $1 |
|
|
@@ -146,7 +147,7 @@ for await (const chunk of stream) {
|
|
|
146
147
|
| `edenai/mistral/codestral-latest` | 256K | | | | | | $0.30 | $0.90 |
|
|
147
148
|
| `edenai/mistral/devstral-2512` | 262K | | | | | | $0.40 | $2 |
|
|
148
149
|
| `edenai/mistral/devstral-medium-latest` | 262K | | | | | | $0.40 | $2 |
|
|
149
|
-
| `edenai/mistral/magistral-medium-latest` | 262K | | | | | | $2 | $
|
|
150
|
+
| `edenai/mistral/magistral-medium-latest` | 262K | | | | | | $2 | $8 |
|
|
150
151
|
| `edenai/mistral/mistral-large-2512` | 262K | | | | | | $0.50 | $2 |
|
|
151
152
|
| `edenai/mistral/mistral-large-latest` | 262K | | | | | | $2 | $6 |
|
|
152
153
|
| `edenai/mistral/mistral-medium-2505` | 131K | | | | | | $0.40 | $2 |
|
|
@@ -4,7 +4,7 @@
|
|
|
4
4
|
|
|
5
5
|
# Fireworks AI
|
|
6
6
|
|
|
7
|
-
Access
|
|
7
|
+
Access 21 Fireworks AI models through Mastra's model router. Authentication is handled automatically using the `FIREWORKS_API_KEY` environment variable.
|
|
8
8
|
|
|
9
9
|
Learn more in the [Fireworks AI documentation](https://fireworks.ai/docs/).
|
|
10
10
|
|
|
@@ -57,6 +57,7 @@ for await (const chunk of stream) {
|
|
|
57
57
|
| `fireworks-ai/accounts/fireworks/models/qwen3p8-2p4t-a95b` | 262K | | | | | | $2 | $6 |
|
|
58
58
|
| `fireworks-ai/accounts/fireworks/models/qwen3p8-max` | 262K | | | | | | $2 | $6 |
|
|
59
59
|
| `fireworks-ai/accounts/fireworks/routers/glm-5p2-fast` | 1.0M | | | | | | $2 | $7 |
|
|
60
|
+
| `fireworks-ai/accounts/fireworks/routers/glm-5p3-fast` | 1.0M | | | | | | $2 | $7 |
|
|
60
61
|
| `fireworks-ai/accounts/fireworks/routers/kimi-k3-fast` | 1.0M | | | | | | $5 | $23 |
|
|
61
62
|
|
|
62
63
|
Model availability, capabilities, context windows, and pricing are sourced from [models.dev](https://models.dev) and may change.
|
|
@@ -42,7 +42,7 @@ for await (const chunk of stream) {
|
|
|
42
42
|
| `hyper/deepseek-v4-flash-0731` | 1.0M | | | | | | $0.44 | $1 |
|
|
43
43
|
| `hyper/deepseek-v4-pro` | 1.0M | | | | | | $2 | $5 |
|
|
44
44
|
| `hyper/deepseek-v4-pro-0813` | 1.0M | | | | | | $1 | $4 |
|
|
45
|
-
| `hyper/gemma-4-26b-a4b-it` | 256K | | | | | | $0.11 | $0.
|
|
45
|
+
| `hyper/gemma-4-26b-a4b-it` | 256K | | | | | | $0.11 | $0.41 |
|
|
46
46
|
| `hyper/glm-5` | 203K | | | | | | $0.86 | $3 |
|
|
47
47
|
| `hyper/glm-5.1` | 203K | | | | | | $1 | $4 |
|
|
48
48
|
| `hyper/glm-5.2` | 1.0M | | | | | | $2 | $5 |
|
|
@@ -51,13 +51,13 @@ for await (const chunk of stream) {
|
|
|
51
51
|
| `hyper/gpt-oss-120b` | 128K | | | | | | $0.19 | $0.70 |
|
|
52
52
|
| `hyper/inkling` | 1.0M | | | | | | $1 | $4 |
|
|
53
53
|
| `hyper/kimi-k2-thinking` | 262K | | | | | | $0.60 | $3 |
|
|
54
|
-
| `hyper/kimi-k2.5` | 262K | | | | | | $0.
|
|
54
|
+
| `hyper/kimi-k2.5` | 262K | | | | | | $0.56 | $3 |
|
|
55
55
|
| `hyper/kimi-k2.6` | 262K | | | | | | $1 | $4 |
|
|
56
56
|
| `hyper/kimi-k2.7-code` | 262K | | | | | | $1 | $4 |
|
|
57
57
|
| `hyper/kimi-k3` | 1.0M | | | | | | $3 | $16 |
|
|
58
58
|
| `hyper/llama-3.3-70b-instruct` | 128K | | | | | | $0.61 | $1 |
|
|
59
59
|
| `hyper/llama-4-maverick-17b-128e-instruct-fp8` | 430K | | | | | | $0.27 | $0.90 |
|
|
60
|
-
| `hyper/minimax-m2.7` | 262K | | | | | | $0.
|
|
60
|
+
| `hyper/minimax-m2.7` | 262K | | | | | | $0.47 | $2 |
|
|
61
61
|
| `hyper/minimax-m3` | 512K | | | | | | $0.33 | $1 |
|
|
62
62
|
| `hyper/qwen3-coder-480b-a35b-instruct-int4-mixed-ar` | 106K | | | | | | $0.45 | $2 |
|
|
63
63
|
| `hyper/qwen3-next-80b-a3b-instruct` | 262K | | | | | | $0.12 | $1 |
|
|
@@ -4,7 +4,7 @@
|
|
|
4
4
|
|
|
5
5
|
# Kilo Gateway
|
|
6
6
|
|
|
7
|
-
Access
|
|
7
|
+
Access 367 Kilo Gateway models through Mastra's model router. Authentication is handled automatically using the `KILO_API_KEY` environment variable.
|
|
8
8
|
|
|
9
9
|
Learn more in the [Kilo Gateway documentation](https://kilo.ai).
|
|
10
10
|
|
|
@@ -42,15 +42,15 @@ for await (const chunk of stream) {
|
|
|
42
42
|
| `kilo/~anthropic/claude-haiku-latest` | 200K | | | | | | $1 | $5 |
|
|
43
43
|
| `kilo/~anthropic/claude-opus-latest` | 1.0M | | | | | | $5 | $25 |
|
|
44
44
|
| `kilo/~anthropic/claude-sonnet-latest` | 1.0M | | | | | | $2 | $10 |
|
|
45
|
-
| `kilo/~deepseek/deepseek-v4-flash-latest` | 1.0M | | | | | | $0.05 | $0.
|
|
45
|
+
| `kilo/~deepseek/deepseek-v4-flash-latest` | 1.0M | | | | | | $0.05 | $0.16 |
|
|
46
46
|
| `kilo/~google/gemini-flash-latest` | 1.0M | | | | | | $0.75 | $4 |
|
|
47
47
|
| `kilo/~google/gemini-pro-latest` | 1.0M | | | | | | $2 | $12 |
|
|
48
|
-
| `kilo/~moonshotai/kimi-latest` | 1.0M | | | | | | $3 | $
|
|
48
|
+
| `kilo/~moonshotai/kimi-latest` | 1.0M | | | | | | $3 | $14 |
|
|
49
49
|
| `kilo/~openai/gpt-latest` | 1.1M | | | | | | $2 | $10 |
|
|
50
50
|
| `kilo/~openai/gpt-mini-latest` | 400K | | | | | | $0.75 | $5 |
|
|
51
51
|
| `kilo/~x-ai/grok-latest` | 500K | | | | | | $2 | $6 |
|
|
52
|
-
| `kilo/~z-ai/glm-flash-latest` | 1.0M | | | | | | $0.07 | $0.
|
|
53
|
-
| `kilo/~z-ai/glm-latest` |
|
|
52
|
+
| `kilo/~z-ai/glm-flash-latest` | 1.0M | | | | | | $0.07 | $0.24 |
|
|
53
|
+
| `kilo/~z-ai/glm-latest` | 1.0M | | | | | | $1 | $4 |
|
|
54
54
|
| `kilo/aion-labs/aion-2.0` | 131K | | | | | | $0.80 | $2 |
|
|
55
55
|
| `kilo/aion-labs/aion-3.0` | 131K | | | | | | $3 | $6 |
|
|
56
56
|
| `kilo/aion-labs/aion-3.0-mini` | 131K | | | | | | $0.70 | $1 |
|
|
@@ -177,9 +177,7 @@ for await (const chunk of stream) {
|
|
|
177
177
|
| `kilo/minimax/minimax-m2.1` | 205K | | | | | | $0.30 | $1 |
|
|
178
178
|
| `kilo/minimax/minimax-m2.5` | 200K | | | | | | $0.30 | $1 |
|
|
179
179
|
| `kilo/minimax/minimax-m2.7` | 205K | | | | | | $0.30 | $1 |
|
|
180
|
-
| `kilo/minimax/minimax-m2.7:free` | 197K | | | | | | — | — |
|
|
181
180
|
| `kilo/minimax/minimax-m3` | 524K | | | | | | $0.30 | $1 |
|
|
182
|
-
| `kilo/minimax/minimax-m3:free` | 1.0M | | | | | | — | — |
|
|
183
181
|
| `kilo/mistralai/codestral-2508` | 256K | | | | | | $0.30 | $0.90 |
|
|
184
182
|
| `kilo/mistralai/devstral-2512` | 262K | | | | | | $0.40 | $2 |
|
|
185
183
|
| `kilo/mistralai/ministral-14b-2512` | 262K | | | | | | $0.20 | $0.20 |
|
|
@@ -265,7 +263,7 @@ for await (const chunk of stream) {
|
|
|
265
263
|
| `kilo/openai/gpt-5.6-luna` | 1.1M | | | | | | $0.20 | $1 |
|
|
266
264
|
| `kilo/openai/gpt-5.6-luna-pro` | 1.1M | | | | | | $0.20 | $1 |
|
|
267
265
|
| `kilo/openai/gpt-5.6-sol` | 1.1M | | | | | | $4 | $20 |
|
|
268
|
-
| `kilo/openai/gpt-5.6-sol-discounted` | 1.1M | | | | | | $
|
|
266
|
+
| `kilo/openai/gpt-5.6-sol-discounted` | 1.1M | | | | | | $2 | $10 |
|
|
269
267
|
| `kilo/openai/gpt-5.6-sol-pro` | 1.1M | | | | | | $4 | $20 |
|
|
270
268
|
| `kilo/openai/gpt-5.6-terra` | 1.1M | | | | | | $2 | $12 |
|
|
271
269
|
| `kilo/openai/gpt-5.6-terra-pro` | 1.1M | | | | | | $2 | $12 |
|
|
@@ -322,7 +320,7 @@ for await (const chunk of stream) {
|
|
|
322
320
|
| `kilo/qwen/qwen3-max` | 262K | | | | | | $0.78 | $4 |
|
|
323
321
|
| `kilo/qwen/qwen3-max-thinking` | 262K | | | | | | $0.78 | $4 |
|
|
324
322
|
| `kilo/qwen/qwen3-next-80b-a3b-instruct` | 262K | | | | | | $0.10 | $0.78 |
|
|
325
|
-
| `kilo/qwen/qwen3-next-80b-a3b-thinking` |
|
|
323
|
+
| `kilo/qwen/qwen3-next-80b-a3b-thinking` | 262K | | | | | | $0.15 | $1 |
|
|
326
324
|
| `kilo/qwen/qwen3-vl-235b-a22b-instruct` | 131K | | | | | | $0.26 | $1 |
|
|
327
325
|
| `kilo/qwen/qwen3-vl-235b-a22b-thinking` | 131K | | | | | | $0.40 | $4 |
|
|
328
326
|
| `kilo/qwen/qwen3-vl-30b-a3b-instruct` | 262K | | | | | | $0.13 | $0.52 |
|
|
@@ -396,7 +394,7 @@ for await (const chunk of stream) {
|
|
|
396
394
|
| `kilo/z-ai/glm-4.5` | 131K | | | | | | $0.60 | $2 |
|
|
397
395
|
| `kilo/z-ai/glm-4.5-air` | 131K | | | | | | $0.13 | $0.85 |
|
|
398
396
|
| `kilo/z-ai/glm-4.5v` | 66K | | | | | | $0.60 | $2 |
|
|
399
|
-
| `kilo/z-ai/glm-4.6` |
|
|
397
|
+
| `kilo/z-ai/glm-4.6` | 198K | | | | | | $0.43 | $2 |
|
|
400
398
|
| `kilo/z-ai/glm-4.6v` | 131K | | | | | | $0.30 | $0.90 |
|
|
401
399
|
| `kilo/z-ai/glm-4.7` | 203K | | | | | | $0.40 | $2 |
|
|
402
400
|
| `kilo/z-ai/glm-4.7-flash` | 203K | | | | | | $0.06 | $0.40 |
|
|
@@ -4,7 +4,7 @@
|
|
|
4
4
|
|
|
5
5
|
# LLM Gateway
|
|
6
6
|
|
|
7
|
-
Access
|
|
7
|
+
Access 373 LLM Gateway models through Mastra's model router. Authentication is handled automatically using the `LLMGATEWAY_API_KEY` environment variable.
|
|
8
8
|
|
|
9
9
|
Learn more in the [LLM Gateway documentation](https://llmgateway.io/docs).
|
|
10
10
|
|
|
@@ -176,7 +176,6 @@ for await (const chunk of stream) {
|
|
|
176
176
|
| `llmgateway-providers/deepinfra/gemma-4-31b-it` | 262K | | | | | | $0.13 | $0.38 |
|
|
177
177
|
| `llmgateway-providers/deepinfra/glm-5.1` | 198K | | | | | | $1 | $4 |
|
|
178
178
|
| `llmgateway-providers/deepinfra/hy3` | 262K | | | | | | $0.14 | $0.58 |
|
|
179
|
-
| `llmgateway-providers/deepinfra/kimi-k2.5` | 256K | | | | | | $0.45 | $2 |
|
|
180
179
|
| `llmgateway-providers/deepinfra/ling-3.0-flash` | 262K | | | | | | $0.06 | $0.18 |
|
|
181
180
|
| `llmgateway-providers/deepinfra/mimo-v2.5` | 262K | | | | | | $0.40 | $2 |
|
|
182
181
|
| `llmgateway-providers/deepinfra/mimo-v2.5-pro` | 1.0M | | | | | | $1 | $3 |
|
|
@@ -4,7 +4,7 @@
|
|
|
4
4
|
|
|
5
5
|
# Moonshot AI (China)
|
|
6
6
|
|
|
7
|
-
Access
|
|
7
|
+
Access 4 Moonshot AI (China) models through Mastra's model router. Authentication is handled automatically using the `MOONSHOT_API_KEY` environment variable.
|
|
8
8
|
|
|
9
9
|
Learn more in the [Moonshot AI (China) documentation](https://platform.moonshot.cn).
|
|
10
10
|
|
|
@@ -19,7 +19,7 @@ const agent = new Agent({
|
|
|
19
19
|
id: "my-agent",
|
|
20
20
|
name: "My Agent",
|
|
21
21
|
instructions: "You are a helpful assistant",
|
|
22
|
-
model: "moonshotai-cn/kimi-k2
|
|
22
|
+
model: "moonshotai-cn/kimi-k2.6"
|
|
23
23
|
});
|
|
24
24
|
|
|
25
25
|
// Generate a response
|
|
@@ -38,12 +38,6 @@ for await (const chunk of stream) {
|
|
|
38
38
|
|
|
39
39
|
| Model | Context | Tools | Reasoning | Image | Audio | Video | Input $/1M | Output $/1M |
|
|
40
40
|
| ---------------------------------------- | ------- | ----- | --------- | ----- | ----- | ----- | ---------- | ----------- |
|
|
41
|
-
| `moonshotai-cn/kimi-k2-0711-preview` | 131K | | | | | | $0.60 | $3 |
|
|
42
|
-
| `moonshotai-cn/kimi-k2-0905-preview` | 262K | | | | | | $0.60 | $3 |
|
|
43
|
-
| `moonshotai-cn/kimi-k2-thinking` | 262K | | | | | | $0.60 | $3 |
|
|
44
|
-
| `moonshotai-cn/kimi-k2-thinking-turbo` | 262K | | | | | | $1 | $8 |
|
|
45
|
-
| `moonshotai-cn/kimi-k2-turbo-preview` | 262K | | | | | | $2 | $10 |
|
|
46
|
-
| `moonshotai-cn/kimi-k2.5` | 262K | | | | | | $0.60 | $3 |
|
|
47
41
|
| `moonshotai-cn/kimi-k2.6` | 262K | | | | | | $0.95 | $4 |
|
|
48
42
|
| `moonshotai-cn/kimi-k2.7-code` | 262K | | | | | | $0.95 | $4 |
|
|
49
43
|
| `moonshotai-cn/kimi-k2.7-code-highspeed` | 262K | | | | | | $2 | $8 |
|
|
@@ -61,7 +55,7 @@ const agent = new Agent({
|
|
|
61
55
|
name: "custom-agent",
|
|
62
56
|
model: {
|
|
63
57
|
url: "https://api.moonshot.cn/anthropic/v1",
|
|
64
|
-
id: "moonshotai-cn/kimi-k2
|
|
58
|
+
id: "moonshotai-cn/kimi-k2.6",
|
|
65
59
|
apiKey: process.env.MOONSHOT_API_KEY,
|
|
66
60
|
headers: {
|
|
67
61
|
"X-Custom-Header": "value"
|
|
@@ -80,7 +74,7 @@ const agent = new Agent({
|
|
|
80
74
|
const useAdvanced = requestContext.task === "complex";
|
|
81
75
|
return useAdvanced
|
|
82
76
|
? "moonshotai-cn/kimi-k3"
|
|
83
|
-
: "moonshotai-cn/kimi-k2
|
|
77
|
+
: "moonshotai-cn/kimi-k2.6";
|
|
84
78
|
}
|
|
85
79
|
});
|
|
86
80
|
```
|
|
@@ -4,7 +4,7 @@
|
|
|
4
4
|
|
|
5
5
|
# Moonshot AI
|
|
6
6
|
|
|
7
|
-
Access
|
|
7
|
+
Access 4 Moonshot AI models through Mastra's model router. Authentication is handled automatically using the `MOONSHOT_API_KEY` environment variable.
|
|
8
8
|
|
|
9
9
|
Learn more in the [Moonshot AI documentation](https://platform.moonshot.ai).
|
|
10
10
|
|
|
@@ -19,7 +19,7 @@ const agent = new Agent({
|
|
|
19
19
|
id: "my-agent",
|
|
20
20
|
name: "My Agent",
|
|
21
21
|
instructions: "You are a helpful assistant",
|
|
22
|
-
model: "moonshotai/kimi-k2
|
|
22
|
+
model: "moonshotai/kimi-k2.6"
|
|
23
23
|
});
|
|
24
24
|
|
|
25
25
|
// Generate a response
|
|
@@ -38,12 +38,6 @@ for await (const chunk of stream) {
|
|
|
38
38
|
|
|
39
39
|
| Model | Context | Tools | Reasoning | Image | Audio | Video | Input $/1M | Output $/1M |
|
|
40
40
|
| ------------------------------------- | ------- | ----- | --------- | ----- | ----- | ----- | ---------- | ----------- |
|
|
41
|
-
| `moonshotai/kimi-k2-0711-preview` | 131K | | | | | | $0.60 | $3 |
|
|
42
|
-
| `moonshotai/kimi-k2-0905-preview` | 262K | | | | | | $0.60 | $3 |
|
|
43
|
-
| `moonshotai/kimi-k2-thinking` | 262K | | | | | | $0.60 | $3 |
|
|
44
|
-
| `moonshotai/kimi-k2-thinking-turbo` | 262K | | | | | | $1 | $8 |
|
|
45
|
-
| `moonshotai/kimi-k2-turbo-preview` | 262K | | | | | | $2 | $10 |
|
|
46
|
-
| `moonshotai/kimi-k2.5` | 262K | | | | | | $0.60 | $3 |
|
|
47
41
|
| `moonshotai/kimi-k2.6` | 262K | | | | | | $0.95 | $4 |
|
|
48
42
|
| `moonshotai/kimi-k2.7-code` | 262K | | | | | | $0.95 | $4 |
|
|
49
43
|
| `moonshotai/kimi-k2.7-code-highspeed` | 262K | | | | | | $2 | $8 |
|
|
@@ -61,7 +55,7 @@ const agent = new Agent({
|
|
|
61
55
|
name: "custom-agent",
|
|
62
56
|
model: {
|
|
63
57
|
url: "https://api.moonshot.ai/anthropic/v1",
|
|
64
|
-
id: "moonshotai/kimi-k2
|
|
58
|
+
id: "moonshotai/kimi-k2.6",
|
|
65
59
|
apiKey: process.env.MOONSHOT_API_KEY,
|
|
66
60
|
headers: {
|
|
67
61
|
"X-Custom-Header": "value"
|
|
@@ -80,7 +74,7 @@ const agent = new Agent({
|
|
|
80
74
|
const useAdvanced = requestContext.task === "complex";
|
|
81
75
|
return useAdvanced
|
|
82
76
|
? "moonshotai/kimi-k3"
|
|
83
|
-
: "moonshotai/kimi-k2
|
|
77
|
+
: "moonshotai/kimi-k2.6";
|
|
84
78
|
}
|
|
85
79
|
});
|
|
86
80
|
```
|
|
@@ -4,7 +4,7 @@
|
|
|
4
4
|
|
|
5
5
|
# NanoGPT
|
|
6
6
|
|
|
7
|
-
Access
|
|
7
|
+
Access 593 NanoGPT models through Mastra's model router. Authentication is handled automatically using the `NANO_GPT_API_KEY` environment variable.
|
|
8
8
|
|
|
9
9
|
Learn more in the [NanoGPT documentation](https://docs.nano-gpt.com).
|
|
10
10
|
|
|
@@ -152,6 +152,7 @@ for await (const chunk of stream) {
|
|
|
152
152
|
| `nano-gpt/deepseek/deepseek-v4-flash-0731:thinking` | 1.0M | | | | | | $0.14 | $0.28 |
|
|
153
153
|
| `nano-gpt/deepseek/deepseek-v4-flash-latest` | 1.0M | | | | | | $0.14 | $0.28 |
|
|
154
154
|
| `nano-gpt/deepseek/deepseek-v4-flash-vision-exp` | 1.0M | | | | | | $0.22 | $0.66 |
|
|
155
|
+
| `nano-gpt/deepseek/deepseek-v4-flash-vision-exp-uncensored` | 524K | | | | | | $0.44 | $2 |
|
|
155
156
|
| `nano-gpt/deepseek/deepseek-v4-flash:thinking` | 1.0M | | | | | | $0.14 | $0.28 |
|
|
156
157
|
| `nano-gpt/deepseek/deepseek-v4-pro` | 1.0M | | | | | | $1 | $2 |
|
|
157
158
|
| `nano-gpt/deepseek/deepseek-v4-pro-0813` | 1.0M | | | | | | $1 | $3 |
|
|
@@ -381,17 +382,17 @@ for await (const chunk of stream) {
|
|
|
381
382
|
| `nano-gpt/openai/gpt-5-nano` | 400K | | | | | | $0.05 | $0.40 |
|
|
382
383
|
| `nano-gpt/openai/gpt-5-pro` | 400K | | | | | | $15 | $120 |
|
|
383
384
|
| `nano-gpt/openai/gpt-5.1` | 400K | | | | | | $1 | $10 |
|
|
384
|
-
| `nano-gpt/openai/gpt-5.1-2025-11-13` |
|
|
385
|
+
| `nano-gpt/openai/gpt-5.1-2025-11-13` | 400K | | | | | | $1 | $10 |
|
|
385
386
|
| `nano-gpt/openai/gpt-5.1-codex` | 400K | | | | | | $1 | $10 |
|
|
386
387
|
| `nano-gpt/openai/gpt-5.1-codex-max` | 400K | | | | | | $3 | $20 |
|
|
387
388
|
| `nano-gpt/openai/gpt-5.1-codex-mini` | 400K | | | | | | $0.25 | $2 |
|
|
388
389
|
| `nano-gpt/openai/gpt-5.2` | 400K | | | | | | $2 | $14 |
|
|
389
390
|
| `nano-gpt/openai/gpt-5.2-codex` | 400K | | | | | | $2 | $14 |
|
|
390
391
|
| `nano-gpt/openai/gpt-5.3-codex` | 400K | | | | | | $2 | $14 |
|
|
391
|
-
| `nano-gpt/openai/gpt-5.4` |
|
|
392
|
+
| `nano-gpt/openai/gpt-5.4` | 1.1M | | | | | | $3 | $15 |
|
|
392
393
|
| `nano-gpt/openai/gpt-5.4-mini` | 400K | | | | | | $0.75 | $5 |
|
|
393
394
|
| `nano-gpt/openai/gpt-5.4-nano` | 400K | | | | | | $0.20 | $1 |
|
|
394
|
-
| `nano-gpt/openai/gpt-5.5` | 1.
|
|
395
|
+
| `nano-gpt/openai/gpt-5.5` | 1.1M | | | | | | $5 | $30 |
|
|
395
396
|
| `nano-gpt/openai/gpt-5.6-luna` | 1.1M | | | | | | $0.20 | $1 |
|
|
396
397
|
| `nano-gpt/openai/gpt-5.6-luna-pro` | 1.1M | | | | | | $0.20 | $1 |
|
|
397
398
|
| `nano-gpt/openai/gpt-5.6-sol` | 1.1M | | | | | | $2 | $10 |
|
|
@@ -426,6 +427,7 @@ for await (const chunk of stream) {
|
|
|
426
427
|
| `nano-gpt/pokee-isaac` | 10.0M | | | | | | $0.15 | $1 |
|
|
427
428
|
| `nano-gpt/poolside/laguna-s-2.1` | 1.0M | | | | | | $0.10 | $0.20 |
|
|
428
429
|
| `nano-gpt/poolside/laguna-s-2.1:thinking` | 1.0M | | | | | | $0.10 | $0.20 |
|
|
430
|
+
| `nano-gpt/poolside/laguna-xs-2.1` | 262K | | | | | | $0.06 | $0.13 |
|
|
429
431
|
| `nano-gpt/qvq-max` | 128K | | | | | | $1 | $5 |
|
|
430
432
|
| `nano-gpt/qwen-3.6-plus` | 992K | | | | | | $0.33 | $2 |
|
|
431
433
|
| `nano-gpt/qwen-long` | 10.0M | | | | | | $0.10 | $0.41 |
|
|
@@ -529,8 +531,8 @@ for await (const chunk of stream) {
|
|
|
529
531
|
| `nano-gpt/TEE/gemma4-31b:thinking` | 262K | | | | | | $0.40 | $1 |
|
|
530
532
|
| `nano-gpt/TEE/glm-5.1` | 203K | | | | | | $2 | $5 |
|
|
531
533
|
| `nano-gpt/TEE/glm-5.1-thinking` | 203K | | | | | | $2 | $5 |
|
|
532
|
-
| `nano-gpt/TEE/glm-5.2` | 1.0M | | | | | | $1 | $
|
|
533
|
-
| `nano-gpt/TEE/glm-5.2:thinking` | 1.0M | | | | | | $1 | $
|
|
534
|
+
| `nano-gpt/TEE/glm-5.2` | 1.0M | | | | | | $1 | $4 |
|
|
535
|
+
| `nano-gpt/TEE/glm-5.2:thinking` | 1.0M | | | | | | $1 | $4 |
|
|
534
536
|
| `nano-gpt/TEE/glm-5.3` | 1.0M | | | | | | $1 | $4 |
|
|
535
537
|
| `nano-gpt/TEE/glm-5.3-flash` | 1.0M | | | | | | $0.15 | $0.50 |
|
|
536
538
|
| `nano-gpt/TEE/gpt-oss-120b` | 131K | | | | | | $2 | $2 |
|
|
@@ -541,7 +543,6 @@ for await (const chunk of stream) {
|
|
|
541
543
|
| `nano-gpt/TEE/llama3-3-70b` | 128K | | | | | | $2 | $3 |
|
|
542
544
|
| `nano-gpt/TEE/muse-glimmer-30b` | 131K | | | | | | $0.35 | $2 |
|
|
543
545
|
| `nano-gpt/TEE/qwen2.5-vl-72b-instruct` | 66K | | | | | | $0.70 | $0.70 |
|
|
544
|
-
| `nano-gpt/TEE/qwen3-8b` | 41K | | | | | | $0.11 | $0.45 |
|
|
545
546
|
| `nano-gpt/TEE/qwen3.5-27b` | 262K | | | | | | $0.30 | $2 |
|
|
546
547
|
| `nano-gpt/TEE/qwen3.5-397b-a17b` | 262K | | | | | | $0.55 | $4 |
|
|
547
548
|
| `nano-gpt/TEE/qwen3.6-27b` | 262K | | | | | | $0.32 | $3 |
|
|
@@ -552,6 +553,7 @@ for await (const chunk of stream) {
|
|
|
552
553
|
| `nano-gpt/tencent/hy4-preview` | 1.0M | | | | | | $0.83 | $3 |
|
|
553
554
|
| `nano-gpt/TheDrummer/Anubis-70B-v1` | 66K | | | | | | $0.31 | $0.31 |
|
|
554
555
|
| `nano-gpt/TheDrummer/Anubis-70B-v1.1` | 32K | | | | | | $0.31 | $0.31 |
|
|
556
|
+
| `nano-gpt/TheDrummer/Artemis-v1.1` | 262K | | | | | | $0.10 | $0.45 |
|
|
555
557
|
| `nano-gpt/TheDrummer/Cydonia-24B-v2` | 33K | | | | | | $0.10 | $0.12 |
|
|
556
558
|
| `nano-gpt/TheDrummer/Cydonia-24B-v4` | 33K | | | | | | $0.20 | $0.24 |
|
|
557
559
|
| `nano-gpt/TheDrummer/Cydonia-24B-v4.1` | 131K | | | | | | $0.35 | $0.55 |
|
|
@@ -624,7 +626,7 @@ for await (const chunk of stream) {
|
|
|
624
626
|
| `nano-gpt/z-ai/glm-5.2:thinking` | 1.0M | | | | | | $0.42 | $1 |
|
|
625
627
|
| `nano-gpt/z-ai/glm-5.3` | 1.0M | | | | | | $1 | $3 |
|
|
626
628
|
| `nano-gpt/z-ai/glm-5.3-flash` | 1.0M | | | | | | $0.07 | $0.25 |
|
|
627
|
-
| `nano-gpt/z-ai/glm-5.3-flash-uncensored` | 1.0M | | | | | | $0.
|
|
629
|
+
| `nano-gpt/z-ai/glm-5.3-flash-uncensored` | 1.0M | | | | | | $0.07 | $0.21 |
|
|
628
630
|
| `nano-gpt/z-ai/glm-5.3:thinking` | 1.0M | | | | | | $1 | $3 |
|
|
629
631
|
| `nano-gpt/z-ai/glm-5v-turbo` | 203K | | | | | | $1 | $4 |
|
|
630
632
|
| `nano-gpt/z-ai/glm-5v-turbo:thinking` | 203K | | | | | | $1 | $4 |
|
|
@@ -4,7 +4,7 @@
|
|
|
4
4
|
|
|
5
5
|
# Ofox
|
|
6
6
|
|
|
7
|
-
Access
|
|
7
|
+
Access 115 Ofox models through Mastra's model router. Authentication is handled automatically using the `OFOX_API_KEY` environment variable.
|
|
8
8
|
|
|
9
9
|
Learn more in the [Ofox documentation](https://ofox.ai/docs).
|
|
10
10
|
|
|
@@ -125,6 +125,7 @@ for await (const chunk of stream) {
|
|
|
125
125
|
| `ofox/openai/gpt-5.6-luna` | 1.1M | | | | | | $0.20 | $1 |
|
|
126
126
|
| `ofox/openai/gpt-5.6-sol` | 1.1M | | | | | | $3 | $15 |
|
|
127
127
|
| `ofox/openai/gpt-5.6-terra` | 1.1M | | | | | | $2 | $12 |
|
|
128
|
+
| `ofox/openai/gpt-6-astra` | 1.1M | | | | | | $10 | $50 |
|
|
128
129
|
| `ofox/volcengine/doubao-seed-1-6` | 256K | | | | | | $0.12 | $0.29 |
|
|
129
130
|
| `ofox/volcengine/doubao-seed-1-6-flash` | 256K | | | | | | $0.03 | $0.22 |
|
|
130
131
|
| `ofox/volcengine/doubao-seed-1-6-vision` | 256K | | | | | | $0.12 | $1 |
|
|
@@ -4,7 +4,7 @@
|
|
|
4
4
|
|
|
5
5
|
# SenseNova (China)
|
|
6
6
|
|
|
7
|
-
Access
|
|
7
|
+
Access 5 SenseNova (China) models through Mastra's model router. Authentication is handled automatically using the `SENSENOVA_API_KEY` environment variable.
|
|
8
8
|
|
|
9
9
|
Learn more in the [SenseNova (China) documentation](https://platform.sensenova.cn/docs).
|
|
10
10
|
|
|
@@ -39,7 +39,9 @@ for await (const chunk of stream) {
|
|
|
39
39
|
| Model | Context | Tools | Reasoning | Image | Audio | Video | Input $/1M | Output $/1M |
|
|
40
40
|
| ------------------------------------ | ------- | ----- | --------- | ----- | ----- | ----- | ---------- | ----------- |
|
|
41
41
|
| `sensenova/deepseek-v4-flash` | 1.0M | | | | | | — | — |
|
|
42
|
+
| `sensenova/deepseek-v4-pro` | 1.0M | | | | | | — | — |
|
|
42
43
|
| `sensenova/glm-5.2` | 1.0M | | | | | | — | — |
|
|
44
|
+
| `sensenova/kimi-k3` | 1.0M | | | | | | — | — |
|
|
43
45
|
| `sensenova/sensenova-6.8-flash-lite` | 262K | | | | | | — | — |
|
|
44
46
|
|
|
45
47
|
Model availability, capabilities, context windows, and pricing are sourced from [models.dev](https://models.dev) and may change.
|
|
@@ -4,7 +4,7 @@
|
|
|
4
4
|
|
|
5
5
|
# Vivgrid
|
|
6
6
|
|
|
7
|
-
Access
|
|
7
|
+
Access 27 Vivgrid models through Mastra's model router. Authentication is handled automatically using the `VIVGRID_API_KEY` environment variable.
|
|
8
8
|
|
|
9
9
|
Learn more in the [Vivgrid documentation](https://docs.vivgrid.com/models).
|
|
10
10
|
|
|
@@ -19,7 +19,7 @@ const agent = new Agent({
|
|
|
19
19
|
id: "my-agent",
|
|
20
20
|
name: "My Agent",
|
|
21
21
|
instructions: "You are a helpful assistant",
|
|
22
|
-
model: "vivgrid/
|
|
22
|
+
model: "vivgrid/claude-fable-5"
|
|
23
23
|
});
|
|
24
24
|
|
|
25
25
|
// Generate a response
|
|
@@ -38,12 +38,16 @@ for await (const chunk of stream) {
|
|
|
38
38
|
|
|
39
39
|
| Model | Context | Tools | Reasoning | Image | Audio | Video | Input $/1M | Output $/1M |
|
|
40
40
|
| --------------------------------------- | ------- | ----- | --------- | ----- | ----- | ----- | ---------- | ----------- |
|
|
41
|
+
| `vivgrid/claude-fable-5` | 1.0M | | | | | | $10 | $50 |
|
|
42
|
+
| `vivgrid/claude-fable-5-1` | 1.0M | | | | | | $10 | $50 |
|
|
41
43
|
| `vivgrid/deepseek-v3.2` | 128K | | | | | | $0.28 | $0.42 |
|
|
42
44
|
| `vivgrid/deepseek-v4-flash` | 1.0M | | | | | | $0.15 | $0.30 |
|
|
43
45
|
| `vivgrid/deepseek-v4-pro` | 1.0M | | | | | | $0.43 | $0.87 |
|
|
46
|
+
| `vivgrid/deepseek-v4-pro-0813` | 1.0M | | | | | | $1 | $3 |
|
|
44
47
|
| `vivgrid/gemini-3.1-flash-lite-preview` | 1.0M | | | | | | $0.25 | $2 |
|
|
45
48
|
| `vivgrid/gemini-3.1-pro-preview` | 1.0M | | | | | | $2 | $12 |
|
|
46
49
|
| `vivgrid/gemini-3.7-flash` | 1.0M | | | | | | $0.75 | $4 |
|
|
50
|
+
| `vivgrid/gemini-3.8-flash` | 1.0M | | | | | | $0.75 | $4 |
|
|
47
51
|
| `vivgrid/glm-5.2` | 1.0M | | | | | | $1 | $4 |
|
|
48
52
|
| `vivgrid/glm-5.3` | 1.0M | | | | | | $1 | $4 |
|
|
49
53
|
| `vivgrid/glm-5.3-flash` | 1.0M | | | | | | $0.15 | $0.50 |
|
|
@@ -59,6 +63,7 @@ for await (const chunk of stream) {
|
|
|
59
63
|
| `vivgrid/gpt-5.6-luna` | 1.1M | | | | | | $1 | $6 |
|
|
60
64
|
| `vivgrid/gpt-5.6-sol` | 1.1M | | | | | | $5 | $30 |
|
|
61
65
|
| `vivgrid/gpt-5.6-terra` | 1.1M | | | | | | $3 | $15 |
|
|
66
|
+
| `vivgrid/gpt-6-astra` | 1.1M | | | | | | $10 | $50 |
|
|
62
67
|
| `vivgrid/kimi-k3` | 1.0M | | | | | | $3 | $15 |
|
|
63
68
|
|
|
64
69
|
Model availability, capabilities, context windows, and pricing are sourced from [models.dev](https://models.dev) and may change.
|
|
@@ -73,7 +78,7 @@ const agent = new Agent({
|
|
|
73
78
|
name: "custom-agent",
|
|
74
79
|
model: {
|
|
75
80
|
url: "https://api.vivgrid.com/v1",
|
|
76
|
-
id: "vivgrid/
|
|
81
|
+
id: "vivgrid/claude-fable-5",
|
|
77
82
|
apiKey: process.env.VIVGRID_API_KEY,
|
|
78
83
|
headers: {
|
|
79
84
|
"X-Custom-Header": "value"
|
|
@@ -92,7 +97,7 @@ const agent = new Agent({
|
|
|
92
97
|
const useAdvanced = requestContext.task === "complex";
|
|
93
98
|
return useAdvanced
|
|
94
99
|
? "vivgrid/kimi-k3"
|
|
95
|
-
: "vivgrid/
|
|
100
|
+
: "vivgrid/claude-fable-5";
|
|
96
101
|
}
|
|
97
102
|
});
|
|
98
103
|
```
|
|
@@ -43,7 +43,7 @@ cleanup()
|
|
|
43
43
|
|
|
44
44
|
### Using the `durable` config flag
|
|
45
45
|
|
|
46
|
-
Set `durable: true` on `AgentConfig` and the agent is automatically wrapped with `createDurableAgent` when it's attached to a `Mastra` instance. Use an object to forward advanced options such as `cache`, `pubsub`, `maxSteps`, or `
|
|
46
|
+
Set `durable: true` on `AgentConfig` and the agent is automatically wrapped with `createDurableAgent` when it's attached to a `Mastra` instance. Use an object to forward advanced options such as `cache`, `pubsub`, `maxSteps`, `cleanupTimeoutMs`, or `shouldCache`.
|
|
47
47
|
|
|
48
48
|
```typescript
|
|
49
49
|
import { Mastra } from '@mastra/core'
|
|
@@ -90,6 +90,8 @@ Returns: `DurableAgent`
|
|
|
90
90
|
|
|
91
91
|
**maxSteps** (`number`): Maximum number of steps for the agentic loop.
|
|
92
92
|
|
|
93
|
+
**shouldCache** (`(topic: string) => boolean`): Per-topic opt-out of the replay cache. Return false to publish a topic straight to the underlying PubSub without recording it; subscribers of that topic receive live events only and cannot resume from an offset. Useful for trading replay for minimum publish latency on hot topics when the cache is remote (for example, cross-region Redis). Run-local topics are always excluded, regardless of this option.
|
|
94
|
+
|
|
93
95
|
## `createEventedAgent(options)`
|
|
94
96
|
|
|
95
97
|
Wraps an `Agent` with fire-and-forget durable execution on the built-in workflow engine. Like `createDurableAgent`, it returns a result you stream from, but the underlying workflow runs non-blocking (via `startAsync`) instead of running to completion before the stream is wired up. Use it when you want the run to progress independently of the caller. It doesn't accept `id` or `name` overrides.
|
|
@@ -112,6 +114,8 @@ Returns: `EventedAgent` (a subclass of `DurableAgent`)
|
|
|
112
114
|
|
|
113
115
|
**maxSteps** (`number`): Maximum number of steps for the agentic loop.
|
|
114
116
|
|
|
117
|
+
**shouldCache** (`(topic: string) => boolean`): Per-topic opt-out of the replay cache. Return false to publish a topic straight to the underlying PubSub without recording it; subscribers of that topic receive live events only and cannot resume from an offset. Useful for trading replay for minimum publish latency on hot topics when the cache is remote (for example, cross-region Redis). Run-local topics are always excluded, regardless of this option.
|
|
118
|
+
|
|
115
119
|
## Constructor parameters
|
|
116
120
|
|
|
117
121
|
The `DurableAgent` class accepts the same options as `createDurableAgent`, plus `cleanupTimeoutMs`. Prefer the factory unless you need to subclass.
|
|
@@ -128,6 +132,8 @@ The `DurableAgent` class accepts the same options as `createDurableAgent`, plus
|
|
|
128
132
|
|
|
129
133
|
**maxSteps** (`number`): Maximum number of steps for the agentic loop.
|
|
130
134
|
|
|
135
|
+
**shouldCache** (`(topic: string) => boolean`): Per-topic opt-out of the replay cache. Return false to publish a topic straight to the underlying PubSub without recording it; subscribers of that topic receive live events only and cannot resume from an offset. Useful for trading replay for minimum publish latency on hot topics when the cache is remote (for example, cross-region Redis). Run-local topics are always excluded, regardless of this option.
|
|
136
|
+
|
|
131
137
|
**cleanupTimeoutMs** (`number`): Grace period in milliseconds before registry entries are cleaned up automatically after a stream finishes or errors. Set to 0 to disable auto-cleanup and require a manual cleanup() call. Auto-cleanup does not fire on suspended events. (Default: `30000`)
|
|
132
138
|
|
|
133
139
|
## Methods
|
|
@@ -23,7 +23,7 @@ Saving changed snapshot fields creates a new latest version. Saving identical sn
|
|
|
23
23
|
|
|
24
24
|
If an active version exists, creating a draft doesn't change the version handling published requests. Publishing updates `activeVersionId`. Restoring a historical version copies its configuration into a new inactive draft.
|
|
25
25
|
|
|
26
|
-
The direct namespace methods and REST APIs
|
|
26
|
+
The direct namespace methods and the REST APIs agree on this. `editor.prompt.update()` and `editor.agent.update()` both create an inactive draft. Neither assigns the new version to `activeVersionId`. To publish a version from the SDK, pass it explicitly: `editor.agent.update({ id, status: 'published', activeVersionId: version.id })`. The stored-agent REST `PATCH` route also creates an inactive draft unless `autoPublish` is enabled, and `POST /stored/agents/:id/versions/:versionId/activate` publishes it.
|
|
27
27
|
|
|
28
28
|
When a generic stored resource has no active version, published resolution can fall back to the latest snapshot. For a code-defined agent override, requesting `status: 'published'` without an active override returns the original code agent.
|
|
29
29
|
|
|
@@ -61,7 +61,7 @@ const graphTool = createGraphRAGTool({
|
|
|
61
61
|
|
|
62
62
|
The tool returns an object with:
|
|
63
63
|
|
|
64
|
-
**relevantContext** (`string`):
|
|
64
|
+
**relevantContext** (`string[]`): Array of chunk text strings for the most relevant document chunks, in rank order, retrieved using graph-based ranking. Text is read from the text field of each chunk's metadata.
|
|
65
65
|
|
|
66
66
|
**sources** (`QueryResult[]`): Array of full retrieval result objects. Each object contains all information needed to reference the original document, chunk, and similarity score.
|
|
67
67
|
|
|
@@ -77,6 +77,8 @@ The tool returns an object with:
|
|
|
77
77
|
}
|
|
78
78
|
```
|
|
79
79
|
|
|
80
|
+
The graph is built from the `text` field of each result's metadata, so `document` (and `relevantContext`) contain that text. Store the chunk text under `metadata.text` when upserting.
|
|
81
|
+
|
|
80
82
|
## Default tool description
|
|
81
83
|
|
|
82
84
|
The default description focuses on:
|
|
@@ -79,7 +79,7 @@ const queryTool = createVectorQueryTool({
|
|
|
79
79
|
|
|
80
80
|
The tool returns an object with:
|
|
81
81
|
|
|
82
|
-
**relevantContext** (`
|
|
82
|
+
**relevantContext** (`any[]`): Array of metadata objects for the most relevant chunks, in rank order (one entry per result). The chunk text is available at relevantContext\[i].text when it was stored in metadata during ingestion.
|
|
83
83
|
|
|
84
84
|
**sources** (`QueryResult[]`): Array of full retrieval result objects. Each object contains all information needed to reference the original document, chunk, and similarity score.
|
|
85
85
|
|
|
@@ -95,6 +95,8 @@ The tool returns an object with:
|
|
|
95
95
|
}
|
|
96
96
|
```
|
|
97
97
|
|
|
98
|
+
`document` is only populated by vector stores whose `query()` returns document content, such as Chroma, Elasticsearch, LanceDB, and MongoDB. For other stores, such as PgVector, it's an empty string. Read the chunk text from `sources[i].metadata.text` (or whichever metadata key you stored it under).
|
|
99
|
+
|
|
98
100
|
## Default tool description
|
|
99
101
|
|
|
100
102
|
The default description focuses on:
|
|
@@ -489,7 +491,7 @@ The tool is created with:
|
|
|
489
491
|
|
|
490
492
|
- **ID**: `VectorQuery {vectorStoreName} {indexName} Tool`
|
|
491
493
|
- **Input Schema**: Requires queryText and filter objects
|
|
492
|
-
- **Output Schema**: Returns relevantContext
|
|
494
|
+
- **Output Schema**: Returns `relevantContext` (array of chunk metadata) and `sources` (array of `QueryResult`)
|
|
493
495
|
|
|
494
496
|
## Related
|
|
495
497
|
|
package/package.json
CHANGED
|
@@ -1,6 +1,6 @@
|
|
|
1
1
|
{
|
|
2
2
|
"name": "@mastra/mcp-docs-server",
|
|
3
|
-
"version": "1.2.24-alpha.
|
|
3
|
+
"version": "1.2.24-alpha.18",
|
|
4
4
|
"description": "MCP server for accessing Mastra.ai documentation, changelogs, and news.",
|
|
5
5
|
"type": "module",
|
|
6
6
|
"main": "dist/index.js",
|
|
@@ -27,8 +27,8 @@
|
|
|
27
27
|
"jsdom": "^26.1.0",
|
|
28
28
|
"local-pkg": "^1.1.2",
|
|
29
29
|
"zod": "^4.4.3",
|
|
30
|
-
"@mastra/
|
|
31
|
-
"@mastra/
|
|
30
|
+
"@mastra/core": "1.65.0-alpha.9",
|
|
31
|
+
"@mastra/mcp": "^1.17.3"
|
|
32
32
|
},
|
|
33
33
|
"devDependencies": {
|
|
34
34
|
"@hono/node-server": "^2.0.0",
|
|
@@ -44,9 +44,9 @@
|
|
|
44
44
|
"tsx": "^4.23.1",
|
|
45
45
|
"typescript": "^7.0.2",
|
|
46
46
|
"vitest": "4.1.10",
|
|
47
|
-
"@internal/types-builder": "0.0.105",
|
|
48
47
|
"@internal/lint": "0.0.130",
|
|
49
|
-
"@
|
|
48
|
+
"@internal/types-builder": "0.0.105",
|
|
49
|
+
"@mastra/core": "1.65.0-alpha.9"
|
|
50
50
|
},
|
|
51
51
|
"homepage": "https://mastra.ai",
|
|
52
52
|
"repository": {
|