@mastra/mcp-docs-server 1.2.17-alpha.18 → 1.2.17-alpha.20

This diff represents the content of publicly available package versions that have been released to one of the supported registries. The information contained in this diff is provided for informational purposes only and reflects changes between package versions as they appear in their respective public registries.
Files changed (32) hide show
  1. package/.docs/docs/deployment/cloud-providers.md +1 -0
  2. package/.docs/docs/deployment/overview.md +1 -0
  3. package/.docs/docs/sandbox/overview.md +1 -0
  4. package/.docs/docs/storage.md +1 -0
  5. package/.docs/docs/workflows/control-flow.md +0 -4
  6. package/.docs/docs/workflows/human-in-the-loop.md +0 -4
  7. package/.docs/docs/workflows/suspend-and-resume.md +0 -4
  8. package/.docs/integrations/deploy/render.md +389 -0
  9. package/.docs/integrations/sandboxes/cloudflare-sandbox.md +118 -0
  10. package/.docs/integrations.md +4 -0
  11. package/.docs/models/environment-variables.md +2 -2
  12. package/.docs/models/gateways/merge-gateway.md +212 -0
  13. package/.docs/models/gateways/openrouter.md +4 -1
  14. package/.docs/models/gateways/vercel.md +2 -1
  15. package/.docs/models/gateways.md +1 -0
  16. package/.docs/models/index.md +1 -1
  17. package/.docs/models/providers/ambient.md +2 -2
  18. package/.docs/models/providers/chutes.md +1 -1
  19. package/.docs/models/providers/edenai.md +5 -5
  20. package/.docs/models/providers/hetzner.md +6 -8
  21. package/.docs/models/providers/hyper.md +5 -5
  22. package/.docs/models/providers/kilo.md +5 -3
  23. package/.docs/models/providers/llmgateway.md +3 -3
  24. package/.docs/models/providers/scx-ai.md +76 -0
  25. package/.docs/models/providers/wandb.md +2 -1
  26. package/.docs/models/providers.md +1 -2
  27. package/.docs/reference/editor/tool-provider.md +107 -0
  28. package/.docs/reference/processors/skill-search-processor.md +2 -0
  29. package/.docs/reference/tools/mcp-server.md +1 -1
  30. package/CHANGELOG.md +14 -0
  31. package/package.json +5 -5
  32. package/.docs/models/providers/merge-gateway.md +0 -268
@@ -0,0 +1,212 @@
1
+ > Discover all available pages from the documentation index: https://mastra.ai/llms.txt
2
+
3
+ # ![Merge Gateway logo](https://models.dev/logos/merge-gateway.svg)Merge Gateway
4
+
5
+ Merge Gateway aggregates models from multiple providers with enhanced features like rate limiting and failover. Access 174 models through Mastra's model router.
6
+
7
+ Learn more in the [Merge Gateway documentation](https://docs.merge.dev/merge-gateway).
8
+
9
+ ## Usage
10
+
11
+ ```typescript
12
+ import { Agent } from "@mastra/core/agent";
13
+
14
+ const agent = new Agent({
15
+ id: "my-agent",
16
+ name: "My Agent",
17
+ instructions: "You are a helpful assistant",
18
+ model: "merge-gateway/anthropic/claude-3-7-sonnet-20250219"
19
+ });
20
+ ```
21
+
22
+ > **Note:** Mastra uses the OpenAI-compatible `/chat/completions` endpoint. Some provider-specific features may not be available. Check the [Merge Gateway documentation](https://docs.merge.dev/merge-gateway) for details.
23
+
24
+ ## Configuration
25
+
26
+ ```bash
27
+ # Use gateway API key
28
+ MERGE_GATEWAY_API_KEY=your-gateway-key
29
+
30
+ # Or use provider API keys directly
31
+ OPENAI_API_KEY=sk-...
32
+ ANTHROPIC_API_KEY=ant-...
33
+ ```
34
+
35
+ ## Available models
36
+
37
+ | Model |
38
+ | ------------------------------------------------ |
39
+ | `anthropic/claude-3-7-sonnet-20250219` |
40
+ | `anthropic/claude-fable-5` |
41
+ | `anthropic/claude-haiku-4-5-20251001` |
42
+ | `anthropic/claude-opus-4-1-20250805` |
43
+ | `anthropic/claude-opus-4-20250514` |
44
+ | `anthropic/claude-opus-4-5-20251101` |
45
+ | `anthropic/claude-opus-4-6` |
46
+ | `anthropic/claude-opus-4-7` |
47
+ | `anthropic/claude-opus-4-8` |
48
+ | `anthropic/claude-opus-5` |
49
+ | `anthropic/claude-sonnet-4-20250514` |
50
+ | `anthropic/claude-sonnet-4-5-20250929` |
51
+ | `anthropic/claude-sonnet-4-6` |
52
+ | `anthropic/claude-sonnet-5` |
53
+ | `bytedance/dola-seed-2.0-code` |
54
+ | `bytedance/dola-seed-2.0-code-preview` |
55
+ | `bytedance/dola-seed-2.0-lite` |
56
+ | `bytedance/dola-seed-2.0-mini` |
57
+ | `bytedance/dola-seed-2.0-pro` |
58
+ | `cohere/command-a-03-2025` |
59
+ | `cohere/command-r-08-2024` |
60
+ | `cohere/command-r-plus-08-2024` |
61
+ | `cohere/command-r7b-12-2024` |
62
+ | `deepseek/deepseek-r1` |
63
+ | `deepseek/deepseek-v3` |
64
+ | `deepseek/deepseek-v3.1` |
65
+ | `deepseek/deepseek-v3.2` |
66
+ | `deepseek/deepseek-v4-flash` |
67
+ | `deepseek/deepseek-v4-flash-0731` |
68
+ | `deepseek/deepseek-v4-pro` |
69
+ | `deepseek/deepseek-v4-pro-0813` |
70
+ | `google/gemini-2.5-computer-use-preview-10-2025` |
71
+ | `google/gemini-2.5-flash` |
72
+ | `google/gemini-2.5-flash-image` |
73
+ | `google/gemini-2.5-flash-lite` |
74
+ | `google/gemini-2.5-pro` |
75
+ | `google/gemini-3-flash-preview` |
76
+ | `google/gemini-3-pro-image` |
77
+ | `google/gemini-3-pro-preview` |
78
+ | `google/gemini-3.1-flash-image` |
79
+ | `google/gemini-3.1-flash-lite` |
80
+ | `google/gemini-3.1-flash-lite-preview` |
81
+ | `google/gemini-3.1-pro-preview` |
82
+ | `google/gemini-3.1-pro-preview-customtools` |
83
+ | `google/gemini-3.5-flash` |
84
+ | `google/gemini-3.5-flash-lite` |
85
+ | `google/gemini-3.6-flash` |
86
+ | `google/gemini-3.7-flash` |
87
+ | `google/gemini-embedding-001` |
88
+ | `google/gemini-flash-latest` |
89
+ | `google/gemini-flash-lite-latest` |
90
+ | `google/gemma-4-26b-a4b-it` |
91
+ | `google/gemma-4-31b-it` |
92
+ | `meta/llama-3.1-8b-instruct` |
93
+ | `meta/llama-3.3-70b-instruct` |
94
+ | `meta/muse-spark-1.1` |
95
+ | `meta/muse-spark-1.2` |
96
+ | `minimax/minimax-m2` |
97
+ | `minimax/minimax-m2.1` |
98
+ | `minimax/minimax-m2.5` |
99
+ | `minimax/minimax-m2.5-highspeed` |
100
+ | `minimax/minimax-m2.7` |
101
+ | `minimax/minimax-m2.7-highspeed` |
102
+ | `minimax/minimax-m3` |
103
+ | `mistral/codestral-latest` |
104
+ | `mistral/devstral-2512` |
105
+ | `mistral/devstral-medium-2507` |
106
+ | `mistral/devstral-medium-latest` |
107
+ | `mistral/devstral-small-2507` |
108
+ | `mistral/magistral-medium-latest` |
109
+ | `mistral/mistral-large-2411` |
110
+ | `mistral/mistral-large-2512` |
111
+ | `mistral/mistral-large-latest` |
112
+ | `mistral/mistral-medium-2505` |
113
+ | `mistral/mistral-medium-latest` |
114
+ | `mistral/mistral-small-latest` |
115
+ | `mistral/pixtral-large-latest` |
116
+ | `moonshot/kimi-k2.5` |
117
+ | `moonshot/kimi-k2.6` |
118
+ | `moonshot/kimi-k2.7-code` |
119
+ | `moonshot/kimi-k2.7-code-highspeed` |
120
+ | `moonshot/kimi-k3` |
121
+ | `moonshotai/kimi-k2-thinking` |
122
+ | `nvidia/nemotron-3.5-lightning-30b-a3b` |
123
+ | `nvidia/nemotron-nano-9b-v2` |
124
+ | `openai/gpt-3.5-turbo` |
125
+ | `openai/gpt-4` |
126
+ | `openai/gpt-4-turbo` |
127
+ | `openai/gpt-4.1` |
128
+ | `openai/gpt-4.1-mini` |
129
+ | `openai/gpt-4.1-nano` |
130
+ | `openai/gpt-4o` |
131
+ | `openai/gpt-4o-2024-05-13` |
132
+ | `openai/gpt-4o-2024-08-06` |
133
+ | `openai/gpt-4o-2024-11-20` |
134
+ | `openai/gpt-4o-mini` |
135
+ | `openai/gpt-5` |
136
+ | `openai/gpt-5-chat-latest` |
137
+ | `openai/gpt-5-mini` |
138
+ | `openai/gpt-5-nano` |
139
+ | `openai/gpt-5.1` |
140
+ | `openai/gpt-5.1-chat-latest` |
141
+ | `openai/gpt-5.2` |
142
+ | `openai/gpt-5.2-chat-latest` |
143
+ | `openai/gpt-5.3-chat-latest` |
144
+ | `openai/gpt-5.4` |
145
+ | `openai/gpt-5.4-mini` |
146
+ | `openai/gpt-5.4-nano` |
147
+ | `openai/gpt-5.5` |
148
+ | `openai/gpt-5.6-luna` |
149
+ | `openai/gpt-5.6-sol` |
150
+ | `openai/gpt-5.6-terra` |
151
+ | `openai/gpt-oss-120b` |
152
+ | `openai/gpt-oss-20b` |
153
+ | `openai/gpt-oss-safeguard-120b` |
154
+ | `openai/o1` |
155
+ | `openai/o3` |
156
+ | `openai/o3-mini` |
157
+ | `openai/o4-mini` |
158
+ | `qwen/qwen-flash` |
159
+ | `qwen/qwen-plus` |
160
+ | `qwen/qwen3-235b-a22b` |
161
+ | `qwen/qwen3-235b-a22b-instruct-2507` |
162
+ | `qwen/qwen3-30b-a3b` |
163
+ | `qwen/qwen3-32b` |
164
+ | `qwen/qwen3-coder-480b-a35b-instruct` |
165
+ | `qwen/qwen3-coder-flash` |
166
+ | `qwen/qwen3-coder-next` |
167
+ | `qwen/qwen3-coder-plus` |
168
+ | `qwen/qwen3-max` |
169
+ | `qwen/qwen3-next-80b-a3b-instruct` |
170
+ | `qwen/qwen3-next-80b-a3b-thinking` |
171
+ | `qwen/qwen3-vl-235b-a22b-instruct` |
172
+ | `qwen/qwen3-vl-235b-a22b-thinking` |
173
+ | `qwen/qwen3-vl-plus` |
174
+ | `qwen/qwen3.5-122b-a10b` |
175
+ | `qwen/qwen3.5-27b` |
176
+ | `qwen/qwen3.5-35b-a3b` |
177
+ | `qwen/qwen3.5-397b-a17b` |
178
+ | `qwen/qwen3.5-9b` |
179
+ | `qwen/qwen3.5-flash` |
180
+ | `qwen/qwen3.5-plus` |
181
+ | `qwen/qwen3.6-27b` |
182
+ | `qwen/qwen3.6-35b-a3b` |
183
+ | `qwen/qwen3.6-flash` |
184
+ | `qwen/qwen3.6-max-preview` |
185
+ | `qwen/qwen3.6-plus` |
186
+ | `qwen/qwen3.7-max` |
187
+ | `qwen/qwen3.7-plus` |
188
+ | `qwen/qwen3.8-2.4t-a95b` |
189
+ | `qwen/qwen3.8-max` |
190
+ | `sakana/fugu-ultra` |
191
+ | `sakana/sakana-namazu` |
192
+ | `thinkingmachines/inkling` |
193
+ | `writer/palmyra-x4` |
194
+ | `writer/palmyra-x5` |
195
+ | `xai/grok-4.20-0309-non-reasoning` |
196
+ | `xai/grok-4.20-0309-reasoning` |
197
+ | `xai/grok-4.3` |
198
+ | `xai/grok-4.5` |
199
+ | `xai/grok-4.6` |
200
+ | `xai/grok-build-0.1` |
201
+ | `zai/glm-4.5` |
202
+ | `zai/glm-4.5-air` |
203
+ | `zai/glm-4.5v` |
204
+ | `zai/glm-4.6` |
205
+ | `zai/glm-4.7` |
206
+ | `zai/glm-4.7-flash` |
207
+ | `zai/glm-4.7-flashx` |
208
+ | `zai/glm-5` |
209
+ | `zai/glm-5-turbo` |
210
+ | `zai/glm-5.1` |
211
+ | `zai/glm-5.2` |
212
+ | `zai/glm-5.3` |
@@ -2,7 +2,7 @@
2
2
 
3
3
  # ![OpenRouter logo](https://models.dev/logos/openrouter.svg)OpenRouter
4
4
 
5
- OpenRouter aggregates models from multiple providers with enhanced features like rate limiting and failover. Access 350 models through Mastra's model router.
5
+ OpenRouter aggregates models from multiple providers with enhanced features like rate limiting and failover. Access 353 models through Mastra's model router.
6
6
 
7
7
  Learn more in the [OpenRouter documentation](https://openrouter.ai/models).
8
8
 
@@ -149,6 +149,7 @@ ANTHROPIC_API_KEY=ant-...
149
149
  | `kwaipilot/kat-coder-air-v2.5` |
150
150
  | `kwaipilot/kat-coder-pro-v2` |
151
151
  | `kwaipilot/kat-coder-pro-v2.5` |
152
+ | `liquid/lfm-2.5-2.6b:free` |
152
153
  | `mancer/weaver` |
153
154
  | `meituan/longcat-2.0` |
154
155
  | `meta-llama/llama-3.1-70b-instruct` |
@@ -385,4 +386,6 @@ ANTHROPIC_API_KEY=ant-...
385
386
  | `z-ai/glm-5-turbo` |
386
387
  | `z-ai/glm-5.1` |
387
388
  | `z-ai/glm-5.2` |
389
+ | `z-ai/glm-5.2:free` |
390
+ | `z-ai/glm-5.3` |
388
391
  | `z-ai/glm-5v-turbo` |
@@ -2,7 +2,7 @@
2
2
 
3
3
  # ![Vercel logo](https://models.dev/logos/vercel.svg)Vercel
4
4
 
5
- Vercel aggregates models from multiple providers with enhanced features like rate limiting and failover. Access 347 models through Mastra's model router.
5
+ Vercel aggregates models from multiple providers with enhanced features like rate limiting and failover. Access 348 models through Mastra's model router.
6
6
 
7
7
  Learn more in the [Vercel documentation](https://ai-sdk.dev/providers/ai-sdk-providers).
8
8
 
@@ -382,4 +382,5 @@ ANTHROPIC_API_KEY=ant-...
382
382
  | `zai/glm-5.1` |
383
383
  | `zai/glm-5.2` |
384
384
  | `zai/glm-5.2-fast` |
385
+ | `zai/glm-5.3` |
385
386
  | `zai/glm-5v-turbo` |
@@ -12,6 +12,7 @@ Create custom gateways for private LLM deployments or specialized provider integ
12
12
 
13
13
  - [Azure OpenAI](https://mastra.ai/models/gateways/azure-openai)
14
14
  - [Mastra](https://mastra.ai/models/gateways/mastra)
15
+ - [Merge Gateway](https://mastra.ai/models/gateways/merge-gateway)
15
16
  - [Neon](https://mastra.ai/models/gateways/neon)
16
17
  - [Netlify](https://mastra.ai/models/gateways/netlify)
17
18
  - [OpenRouter](https://mastra.ai/models/gateways/openrouter)
@@ -2,7 +2,7 @@
2
2
 
3
3
  # Model Providers
4
4
 
5
- Mastra provides a unified interface for working with LLMs across multiple providers, giving you access to 6097 models from 178 providers through a single API.
5
+ Mastra provides a unified interface for working with LLMs across multiple providers, giving you access to 6106 models from 178 providers through a single API.
6
6
 
7
7
  ## Features
8
8
 
@@ -36,14 +36,14 @@ for await (const chunk of stream) {
36
36
 
37
37
  | Model | Context | Tools | Reasoning | Image | Audio | Video | Input $/1M | Output $/1M |
38
38
  | ----------------------------------------- | ------- | ----- | --------- | ----- | ----- | ----- | ---------- | ----------- |
39
- | `ambient/ambient/large` | 203K | | | | | | $1 | $4 |
39
+ | `ambient/ambient/large` | 203K | | | | | | $0.60 | $2 |
40
40
  | `ambient/deepseek/deepseek-v4-flash` | 1.0M | | | | | | $0.14 | $0.28 |
41
41
  | `ambient/deepseek/deepseek-v4-flash-0731` | 1.0M | | | | | | $0.14 | $0.28 |
42
42
  | `ambient/moonshotai/kimi-k2.6` | 262K | | | | | | $0.95 | $4 |
43
43
  | `ambient/moonshotai/kimi-k2.7-code` | 262K | | | | | | $0.69 | $3 |
44
44
  | `ambient/stepfun/step-3.7-flash` | 262K | | | | | | $0.19 | $1 |
45
45
  | `ambient/xiaomi/mimo-v2.5` | 1.0M | | | | | | $0.40 | $2 |
46
- | `ambient/z-ai/glm-5.2` | 203K | | | | | | $1 | $4 |
46
+ | `ambient/z-ai/glm-5.2` | 203K | | | | | | $0.60 | $2 |
47
47
  | `ambient/zai-org/GLM-5.1-FP8` | 203K | | | | | | $1 | $4 |
48
48
  | `ambient/zai-org/GLM-5.2-FP8` | 203K | | | | | | $1 | $4 |
49
49
 
@@ -46,7 +46,7 @@ for await (const chunk of stream) {
46
46
  | `chutes/Qwen/Qwen3-32B-TEE` | 41K | | | | | | $0.10 | $0.42 |
47
47
  | `chutes/Qwen/Qwen3.5-397B-A17B-TEE` | 262K | | | | | | $0.45 | $3 |
48
48
  | `chutes/Qwen/Qwen3.6-27B-TEE` | 262K | | | | | | $0.30 | $2 |
49
- | `chutes/Qwen/Qwen3.8-27B-TEE` | 262K | | | | | | $0.40 | $3 |
49
+ | `chutes/Qwen/Qwen3.8-27B-TEE` | 262K | | | | | | $0.45 | $3 |
50
50
  | `chutes/unsloth/Mistral-Nemo-Instruct-2407-TEE` | 131K | | | | | | $0.02 | $0.10 |
51
51
  | `chutes/zai-org/GLM-5.1-TEE` | 203K | | | | | | $0.98 | $3 |
52
52
  | `chutes/zai-org/GLM-5.2-TEE` | 1.0M | | | | | | $1 | $4 |
@@ -100,8 +100,8 @@ for await (const chunk of stream) {
100
100
  | `edenai/deepinfra/zai-org/GLM-4.7-Flash` | 203K | | | | | | $0.06 | $0.40 |
101
101
  | `edenai/deepseek/deepseek-chat` | 131K | | | | | | $0.28 | $0.42 |
102
102
  | `edenai/deepseek/deepseek-reasoner` | 131K | | | | | | $0.28 | $0.42 |
103
- | `edenai/deepseek/deepseek-v4-flash` | 1.0M | | | | | | $0.22 | $0.66 |
104
- | `edenai/deepseek/deepseek-v4-pro` | 1.0M | | | | | | $0.66 | $2 |
103
+ | `edenai/deepseek/deepseek-v4-flash` | 1.0M | | | | | | $0.44 | $1 |
104
+ | `edenai/deepseek/deepseek-v4-pro` | 1.0M | | | | | | $1 | $4 |
105
105
  | `edenai/fireworks_ai/accounts/fireworks/models/deepseek-v4-flash-0731` | 1.0M | | | | | | $0.14 | $0.28 |
106
106
  | `edenai/fireworks_ai/accounts/fireworks/models/deepseek-v4-pro-0813` | 1.0M | | | | | | $1 | $4 |
107
107
  | `edenai/fireworks_ai/accounts/fireworks/models/gpt-oss-120b` | 131K | | | | | | $0.15 | $0.60 |
@@ -158,7 +158,7 @@ for await (const chunk of stream) {
158
158
  | `edenai/moonshot/kimi-k3` | 1.0M | | | | | | $3 | $15 |
159
159
  | `edenai/nebius/meta-llama/Llama-3.3-70B-Instruct` | 131K | | | | | | $0.13 | $0.40 |
160
160
  | `edenai/nebius/nvidia/nemotron-3-super-120b-a12b` | 8K | | | | | | $0.30 | $0.90 |
161
- | `edenai/nebius/nvidia/Nemotron-3-Ultra-550b-a55b` | 8K | | | | | | $1 | $3 |
161
+ | `edenai/nebius/nvidia/Nemotron-3-Ultra-550b-a55b` | 1.0M | | | | | | $1 | $3 |
162
162
  | `edenai/nebius/openai/gpt-oss-120b` | 131K | | | | | | $0.15 | $0.60 |
163
163
  | `edenai/openai/gpt-3.5-turbo` | 16K | | | | | | $0.50 | $2 |
164
164
  | `edenai/openai/gpt-4` | 8K | | | | | | $30 | $60 |
@@ -203,7 +203,7 @@ for await (const chunk of stream) {
203
203
  | `edenai/perplexityai/sonar-pro` | 200K | | | | | | $3 | $15 |
204
204
  | `edenai/perplexityai/sonar-reasoning-pro` | 128K | | | | | | $2 | $8 |
205
205
  | `edenai/qwen/deepseek-v4-flash-0731` | 1.0M | | | | | | $0.20 | $0.40 |
206
- | `edenai/qwen/deepseek-v4-pro-0813` | 1.0M | | | | | | $1 | $4 |
206
+ | `edenai/qwen/deepseek-v4-pro-0813` | 1.0M | | | | | | $0.66 | $2 |
207
207
  | `edenai/qwen/qwen-max` | 33K | | | | | | $2 | $6 |
208
208
  | `edenai/qwen/qwen-vl-max` | 131K | | | | | | $0.80 | $3 |
209
209
  | `edenai/qwen/qwen-vl-plus` | 131K | | | | | | $0.21 | $0.63 |
@@ -222,7 +222,7 @@ for await (const chunk of stream) {
222
222
  | `edenai/qwen/qwen3.8-max` | 1.0M | | | | | | $2 | $6 |
223
223
  | `edenai/qwen/qwq-plus` | 131K | | | | | | $0.80 | $2 |
224
224
  | `edenai/scaleway/deepseek-v4-flash-0731` | 256K | | | | | | $0.46 | $0.93 |
225
- | `edenai/scaleway/gpt-oss-120b` | 128K | | | | | | $0.17 | $0.70 |
225
+ | `edenai/scaleway/gpt-oss-120b` | 128K | | | | | | $0.17 | $0.69 |
226
226
  | `edenai/scaleway/llama-3.3-70b-instruct` | 128K | | | | | | $1 | $1 |
227
227
  | `edenai/tensorx/deepseek/deepseek-v4-flash-0731` | 1.0M | | | | | | $0.25 | $0.30 |
228
228
  | `edenai/tensorx/moonshotai/kimi-k2.5` | 262K | | | | | | $0.50 | $3 |
@@ -2,7 +2,7 @@
2
2
 
3
3
  # ![Hetzner logo](https://models.dev/logos/hetzner.svg)Hetzner
4
4
 
5
- Access 4 Hetzner models through Mastra's model router. Authentication is handled automatically using the `HETZNER_API_KEY` environment variable.
5
+ Access 2 Hetzner models through Mastra's model router. Authentication is handled automatically using the `HETZNER_API_KEY` environment variable.
6
6
 
7
7
  Learn more in the [Hetzner documentation](https://experiments.hetzner.com).
8
8
 
@@ -17,7 +17,7 @@ const agent = new Agent({
17
17
  id: "my-agent",
18
18
  name: "My Agent",
19
19
  instructions: "You are a helpful assistant",
20
- model: "hetzner/DeepSeek-V4-Flash-0731"
20
+ model: "hetzner/Qwen/Qwen3.6-35B-A3B-FP8"
21
21
  });
22
22
 
23
23
  // Generate a response
@@ -36,10 +36,8 @@ for await (const chunk of stream) {
36
36
 
37
37
  | Model | Context | Tools | Reasoning | Image | Audio | Video | Input $/1M | Output $/1M |
38
38
  | ---------------------------------- | ------- | ----- | --------- | ----- | ----- | ----- | ---------- | ----------- |
39
- | `hetzner/DeepSeek-V4-Flash-0731` | 512K | | | | | | — | — |
40
- | `hetzner/GLM-5.2-NVFP4` | 512K | | | | | | — | — |
41
- | `hetzner/Kimi-K2.7-Code` | 262K | | | | | | — | — |
42
39
  | `hetzner/Qwen/Qwen3.6-35B-A3B-FP8` | 262K | | | | | | — | — |
40
+ | `hetzner/Qwen3.8-27B` | 262K | | | | | | — | — |
43
41
 
44
42
  ## Advanced configuration
45
43
 
@@ -51,7 +49,7 @@ const agent = new Agent({
51
49
  name: "custom-agent",
52
50
  model: {
53
51
  url: "https://inference.hetzner.com/api/v1",
54
- id: "hetzner/DeepSeek-V4-Flash-0731",
52
+ id: "hetzner/Qwen/Qwen3.6-35B-A3B-FP8",
55
53
  apiKey: process.env.HETZNER_API_KEY,
56
54
  headers: {
57
55
  "X-Custom-Header": "value"
@@ -69,8 +67,8 @@ const agent = new Agent({
69
67
  model: ({ requestContext }) => {
70
68
  const useAdvanced = requestContext.task === "complex";
71
69
  return useAdvanced
72
- ? "hetzner/Qwen/Qwen3.6-35B-A3B-FP8"
73
- : "hetzner/DeepSeek-V4-Flash-0731";
70
+ ? "hetzner/Qwen3.8-27B"
71
+ : "hetzner/Qwen/Qwen3.6-35B-A3B-FP8";
74
72
  }
75
73
  });
76
74
  ```
@@ -41,17 +41,17 @@ for await (const chunk of stream) {
41
41
  | `hyper/deepseek-v4-pro` | 1.0M | | | | | | $2 | $5 |
42
42
  | `hyper/deepseek-v4-pro-0813` | 1.0M | | | | | | $1 | $4 |
43
43
  | `hyper/gemma-4-26b-a4b-it` | 256K | | | | | | $0.12 | $0.42 |
44
- | `hyper/glm-5` | 203K | | | | | | $0.80 | $3 |
45
- | `hyper/glm-5.1` | 203K | | | | | | $2 | $5 |
44
+ | `hyper/glm-5` | 203K | | | | | | $0.84 | $3 |
45
+ | `hyper/glm-5.1` | 203K | | | | | | $1 | $4 |
46
46
  | `hyper/glm-5.2` | 1.0M | | | | | | $1 | $4 |
47
47
  | `hyper/gpt-oss-120b` | 131K | | | | | | $0.16 | $0.65 |
48
- | `hyper/kimi-k2.5` | 262K | | | | | | $0.51 | $3 |
48
+ | `hyper/kimi-k2.5` | 262K | | | | | | $0.57 | $3 |
49
49
  | `hyper/kimi-k2.6` | 262K | | | | | | $0.95 | $4 |
50
50
  | `hyper/kimi-k2.7-code` | 256K | | | | | | $0.95 | $4 |
51
51
  | `hyper/kimi-k3` | 1.0M | | | | | | $3 | $15 |
52
- | `hyper/llama-3.3-70b-instruct` | 128K | | | | | | $0.60 | $0.74 |
52
+ | `hyper/llama-3.3-70b-instruct` | 128K | | | | | | $0.64 | $0.77 |
53
53
  | `hyper/llama-4-maverick-17b-128e-instruct-fp8` | 430K | | | | | | $0.27 | $0.90 |
54
- | `hyper/minimax-m2.7` | 262K | | | | | | $0.41 | $2 |
54
+ | `hyper/minimax-m2.7` | 262K | | | | | | $0.47 | $2 |
55
55
  | `hyper/minimax-m3` | 512K | | | | | | $0.33 | $1 |
56
56
  | `hyper/qwen3-coder-480b-a35b-instruct-int4-mixed-ar` | 106K | | | | | | $0.45 | $2 |
57
57
  | `hyper/qwen3-next-80b-a3b-instruct` | 262K | | | | | | $0.12 | $1 |
@@ -2,7 +2,7 @@
2
2
 
3
3
  # ![Kilo Gateway logo](https://models.dev/logos/kilo.svg)Kilo Gateway
4
4
 
5
- Access 358 Kilo Gateway models through Mastra's model router. Authentication is handled automatically using the `KILO_API_KEY` environment variable.
5
+ Access 360 Kilo Gateway models through Mastra's model router. Authentication is handled automatically using the `KILO_API_KEY` environment variable.
6
6
 
7
7
  Learn more in the [Kilo Gateway documentation](https://kilo.ai).
8
8
 
@@ -40,7 +40,7 @@ for await (const chunk of stream) {
40
40
  | `kilo/~anthropic/claude-haiku-latest` | 200K | | | | | | $1 | $5 |
41
41
  | `kilo/~anthropic/claude-opus-latest` | 1.0M | | | | | | $5 | $25 |
42
42
  | `kilo/~anthropic/claude-sonnet-latest` | 1.0M | | | | | | $2 | $10 |
43
- | `kilo/~deepseek/deepseek-v4-flash-latest` | 1.0M | | | | | | $0.08 | $0.16 |
43
+ | `kilo/~deepseek/deepseek-v4-flash-latest` | 262K | | | | | | $0.08 | $0.15 |
44
44
  | `kilo/~google/gemini-flash-latest` | 1.0M | | | | | | $0.38 | $2 |
45
45
  | `kilo/~google/gemini-pro-latest` | 1.0M | | | | | | $2 | $12 |
46
46
  | `kilo/~moonshotai/kimi-latest` | 975K | | | | | | $3 | $13 |
@@ -128,7 +128,7 @@ for await (const chunk of stream) {
128
128
  | `kilo/google/gemini-3.5-flash` | 1.0M | | | | | | $2 | $9 |
129
129
  | `kilo/google/gemini-3.5-flash-lite` | 1.0M | | | | | | $0.30 | $3 |
130
130
  | `kilo/google/gemini-3.6-flash` | 1.0M | | | | | | $0.75 | $4 |
131
- | `kilo/google/gemini-3.7-flash` | 1.0M | | | | | | $0.75 | $4 |
131
+ | `kilo/google/gemini-3.7-flash` | 1.0M | | | | | | $2 | $8 |
132
132
  | `kilo/google/gemma-2-27b-it` | 8K | | | | | | $0.65 | $0.65 |
133
133
  | `kilo/google/gemma-3-12b-it` | 131K | | | | | | $0.05 | $0.15 |
134
134
  | `kilo/google/gemma-3-27b-it` | 131K | | | | | | $0.08 | $0.16 |
@@ -154,6 +154,7 @@ for await (const chunk of stream) {
154
154
  | `kilo/kwaipilot/kat-coder-air-v2.5` | 256K | | | | | | $0.15 | $0.60 |
155
155
  | `kilo/kwaipilot/kat-coder-pro-v2` | 256K | | | | | | $0.30 | $1 |
156
156
  | `kilo/kwaipilot/kat-coder-pro-v2.5` | 256K | | | | | | $0.74 | $3 |
157
+ | `kilo/liquid/lfm-2.5-2.6b:free` | 128K | | | | | | — | — |
157
158
  | `kilo/mancer/weaver` | 8K | | | | | | $0.50 | $0.75 |
158
159
  | `kilo/meituan/longcat-2.0` | 1.0M | | | | | | $0.75 | $3 |
159
160
  | `kilo/meta-llama/llama-3.1-70b-instruct` | 131K | | | | | | $0.40 | $0.40 |
@@ -393,6 +394,7 @@ for await (const chunk of stream) {
393
394
  | `kilo/z-ai/glm-5-turbo` | 203K | | | | | | $1 | $4 |
394
395
  | `kilo/z-ai/glm-5.1` | 200K | | | | | | $1 | $4 |
395
396
  | `kilo/z-ai/glm-5.2` | 1.0M | | | | | | $1 | $4 |
397
+ | `kilo/z-ai/glm-5.3` | 1.0M | | | | | | $1 | $4 |
396
398
  | `kilo/z-ai/glm-5v-turbo` | 203K | | | | | | $1 | $4 |
397
399
 
398
400
  ## Advanced configuration
@@ -55,7 +55,7 @@ for await (const chunk of stream) {
55
55
  | `llmgateway/cosmos3-super-reasoner` | 262K | | | | | | $0.10 | $0.30 |
56
56
  | `llmgateway/custom` | 128K | | | | | | — | — |
57
57
  | `llmgateway/deepseek-v3.2` | 164K | | | | | | $0.26 | $0.38 |
58
- | `llmgateway/deepseek-v4-flash` | 1.1M | | | | | | $0.08 | $0.15 |
58
+ | `llmgateway/deepseek-v4-flash` | 1.1M | | | | | | $0.05 | $0.09 |
59
59
  | `llmgateway/deepseek-v4-pro` | 1.1M | | | | | | $0.43 | $0.87 |
60
60
  | `llmgateway/ernie-4.5-vl-424b-a47b` | 123K | | | | | | $0.42 | $1 |
61
61
  | `llmgateway/fugu-ultra` | 1.0M | | | | | | $5 | $30 |
@@ -176,7 +176,7 @@ for await (const chunk of stream) {
176
176
  | `llmgateway/nemotron-3-nano-30b` | 262K | | | | | | $0.06 | $0.24 |
177
177
  | `llmgateway/nemotron-3-nano-omni` | 262K | | | | | | $0.06 | $0.24 |
178
178
  | `llmgateway/nemotron-3-super-120b` | 262K | | | | | | $0.30 | $0.90 |
179
- | `llmgateway/nemotron-3-ultra-550b` | 1.0M | | | | | | $0.50 | $3 |
179
+ | `llmgateway/nemotron-3-ultra-550b` | 1.0M | | | | | | $0.50 | $2 |
180
180
  | `llmgateway/o1` | 200K | | | | | | $15 | $60 |
181
181
  | `llmgateway/o3` | 200K | | | | | | $2 | $8 |
182
182
  | `llmgateway/o3-mini` | 200K | | | | | | $1 | $4 |
@@ -194,7 +194,7 @@ for await (const chunk of stream) {
194
194
  | `llmgateway/qwen3-30b-a3b-instruct-2507` | 262K | | | | | | $0.10 | $0.30 |
195
195
  | `llmgateway/qwen3-32b` | 41K | | | | | | $0.10 | $0.30 |
196
196
  | `llmgateway/qwen3-coder-30b-a3b-instruct` | 262K | | | | | | $0.07 | $0.27 |
197
- | `llmgateway/qwen3-coder-480b-a35b-instruct` | 262K | | | | | | $0.30 | $1 |
197
+ | `llmgateway/qwen3-coder-480b-a35b-instruct` | 262K | | | | | | $0.38 | $2 |
198
198
  | `llmgateway/qwen3-coder-flash` | 1.0M | | | | | | $0.30 | $2 |
199
199
  | `llmgateway/qwen3-coder-next` | 262K | | | | | | $0.11 | $0.68 |
200
200
  | `llmgateway/qwen3-coder-plus` | 1.0M | | | | | | $6 | $60 |
@@ -0,0 +1,76 @@
1
+ > Discover all available pages from the documentation index: https://mastra.ai/llms.txt
2
+
3
+ # ![SCX.ai logo](https://models.dev/logos/scx-ai.svg)SCX.ai
4
+
5
+ Access 4 SCX.ai models through Mastra's model router. Authentication is handled automatically using the `SCX_API_KEY` environment variable.
6
+
7
+ Learn more in the [SCX.ai documentation](https://platform.scx.ai/docs).
8
+
9
+ ```bash
10
+ SCX_API_KEY=your-api-key
11
+ ```
12
+
13
+ ```typescript
14
+ import { Agent } from "@mastra/core/agent";
15
+
16
+ const agent = new Agent({
17
+ id: "my-agent",
18
+ name: "My Agent",
19
+ instructions: "You are a helpful assistant",
20
+ model: "scx-ai/GLM-5.2"
21
+ });
22
+
23
+ // Generate a response
24
+ const response = await agent.generate("Hello!");
25
+
26
+ // Stream a response
27
+ const stream = await agent.stream("Tell me a story");
28
+ for await (const chunk of stream) {
29
+ console.log(chunk);
30
+ }
31
+ ```
32
+
33
+ > **Note:** Mastra uses the OpenAI-compatible `/chat/completions` endpoint. Some provider-specific features may not be available. Check the [SCX.ai documentation](https://platform.scx.ai/docs) for details.
34
+
35
+ ## Models
36
+
37
+ | Model | Context | Tools | Reasoning | Image | Audio | Video | Input $/1M | Output $/1M |
38
+ | --------------------- | ------- | ----- | --------- | ----- | ----- | ----- | ---------- | ----------- |
39
+ | `scx-ai/GLM-5.2` | 1.0M | | | | | | $0.55 | $2 |
40
+ | `scx-ai/gpt-oss-120b` | 131K | | | | | | $0.17 | $0.55 |
41
+ | `scx-ai/MiniMax-M2.7` | 197K | | | | | | $0.48 | $2 |
42
+ | `scx-ai/Qwen3.8-Max` | 1.0M | | | | | | $2 | $5 |
43
+
44
+ ## Advanced configuration
45
+
46
+ ### Custom headers
47
+
48
+ ```typescript
49
+ const agent = new Agent({
50
+ id: "custom-agent",
51
+ name: "custom-agent",
52
+ model: {
53
+ url: "https://api.scx.ai/v1",
54
+ id: "scx-ai/GLM-5.2",
55
+ apiKey: process.env.SCX_API_KEY,
56
+ headers: {
57
+ "X-Custom-Header": "value"
58
+ }
59
+ }
60
+ });
61
+ ```
62
+
63
+ ### Dynamic model selection
64
+
65
+ ```typescript
66
+ const agent = new Agent({
67
+ id: "dynamic-agent",
68
+ name: "Dynamic Agent",
69
+ model: ({ requestContext }) => {
70
+ const useAdvanced = requestContext.task === "complex";
71
+ return useAdvanced
72
+ ? "scx-ai/gpt-oss-120b"
73
+ : "scx-ai/GLM-5.2";
74
+ }
75
+ });
76
+ ```
@@ -2,7 +2,7 @@
2
2
 
3
3
  # ![Weights & Biases logo](https://models.dev/logos/wandb.svg)Weights & Biases
4
4
 
5
- Access 28 Weights & Biases models through Mastra's model router. Authentication is handled automatically using the `WANDB_API_KEY` environment variable.
5
+ Access 29 Weights & Biases models through Mastra's model router. Authentication is handled automatically using the `WANDB_API_KEY` environment variable.
6
6
 
7
7
  Learn more in the [Weights & Biases documentation](https://docs.wandb.ai).
8
8
 
@@ -62,6 +62,7 @@ for await (const chunk of stream) {
62
62
  | `wandb/Qwen/Qwen3.5-35B-A3B` | 262K | | | | | | $0.25 | $1 |
63
63
  | `wandb/Qwen/Qwen3.6-27B` | 262K | | | | | | $0.60 | $4 |
64
64
  | `wandb/Qwen/Qwen3.6-35B-A3B` | 262K | | | | | | $0.25 | $1 |
65
+ | `wandb/Qwen/Qwen3.8-27B` | 262K | | | | | | $0.40 | $3 |
65
66
  | `wandb/zai-org/GLM-5.1` | 203K | | | | | | $1 | $4 |
66
67
  | `wandb/zai-org/GLM-5.2` | 262K | | | | | | $0.76 | $2 |
67
68
 
@@ -91,7 +91,6 @@ Direct access to individual AI model providers. Each provider offers unique mode
91
91
  - [LucidQuery](https://mastra.ai/models/providers/lucidquery)
92
92
  - [Lynkr](https://mastra.ai/models/providers/lynkr)
93
93
  - [Meganova](https://mastra.ai/models/providers/meganova)
94
- - [Merge Gateway](https://mastra.ai/models/providers/merge-gateway)
95
94
  - [Meta](https://mastra.ai/models/providers/meta)
96
95
  - [MiniMax (minimax.io)](https://mastra.ai/models/providers/minimax)
97
96
  - [MiniMax (minimaxi.com)](https://mastra.ai/models/providers/minimax-cn)
@@ -135,7 +134,7 @@ Direct access to individual AI model providers. Each provider offers unique mode
135
134
  - [Sarvam AI](https://mastra.ai/models/providers/sarvam)
136
135
  - [Scaleway](https://mastra.ai/models/providers/scaleway)
137
136
  - [SCNet Token Plan](https://mastra.ai/models/providers/scnet-token-plan)
138
- - [SCX.ai](https://mastra.ai/models/providers/scx)
137
+ - [SCX.ai](https://mastra.ai/models/providers/scx-ai)
139
138
  - [SiliconFlow](https://mastra.ai/models/providers/siliconflow)
140
139
  - [SiliconFlow (China)](https://mastra.ai/models/providers/siliconflow-cn)
141
140
  - [Snowflake Cortex](https://mastra.ai/models/providers/snowflake-cortex)