@mastra/mcp-docs-server 1.2.17-alpha.18 → 1.2.17-alpha.20
This diff represents the content of publicly available package versions that have been released to one of the supported registries. The information contained in this diff is provided for informational purposes only and reflects changes between package versions as they appear in their respective public registries.
- package/.docs/docs/deployment/cloud-providers.md +1 -0
- package/.docs/docs/deployment/overview.md +1 -0
- package/.docs/docs/sandbox/overview.md +1 -0
- package/.docs/docs/storage.md +1 -0
- package/.docs/docs/workflows/control-flow.md +0 -4
- package/.docs/docs/workflows/human-in-the-loop.md +0 -4
- package/.docs/docs/workflows/suspend-and-resume.md +0 -4
- package/.docs/integrations/deploy/render.md +389 -0
- package/.docs/integrations/sandboxes/cloudflare-sandbox.md +118 -0
- package/.docs/integrations.md +4 -0
- package/.docs/models/environment-variables.md +2 -2
- package/.docs/models/gateways/merge-gateway.md +212 -0
- package/.docs/models/gateways/openrouter.md +4 -1
- package/.docs/models/gateways/vercel.md +2 -1
- package/.docs/models/gateways.md +1 -0
- package/.docs/models/index.md +1 -1
- package/.docs/models/providers/ambient.md +2 -2
- package/.docs/models/providers/chutes.md +1 -1
- package/.docs/models/providers/edenai.md +5 -5
- package/.docs/models/providers/hetzner.md +6 -8
- package/.docs/models/providers/hyper.md +5 -5
- package/.docs/models/providers/kilo.md +5 -3
- package/.docs/models/providers/llmgateway.md +3 -3
- package/.docs/models/providers/scx-ai.md +76 -0
- package/.docs/models/providers/wandb.md +2 -1
- package/.docs/models/providers.md +1 -2
- package/.docs/reference/editor/tool-provider.md +107 -0
- package/.docs/reference/processors/skill-search-processor.md +2 -0
- package/.docs/reference/tools/mcp-server.md +1 -1
- package/CHANGELOG.md +14 -0
- package/package.json +5 -5
- package/.docs/models/providers/merge-gateway.md +0 -268
|
@@ -0,0 +1,212 @@
|
|
|
1
|
+
> Discover all available pages from the documentation index: https://mastra.ai/llms.txt
|
|
2
|
+
|
|
3
|
+
# Merge Gateway
|
|
4
|
+
|
|
5
|
+
Merge Gateway aggregates models from multiple providers with enhanced features like rate limiting and failover. Access 174 models through Mastra's model router.
|
|
6
|
+
|
|
7
|
+
Learn more in the [Merge Gateway documentation](https://docs.merge.dev/merge-gateway).
|
|
8
|
+
|
|
9
|
+
## Usage
|
|
10
|
+
|
|
11
|
+
```typescript
|
|
12
|
+
import { Agent } from "@mastra/core/agent";
|
|
13
|
+
|
|
14
|
+
const agent = new Agent({
|
|
15
|
+
id: "my-agent",
|
|
16
|
+
name: "My Agent",
|
|
17
|
+
instructions: "You are a helpful assistant",
|
|
18
|
+
model: "merge-gateway/anthropic/claude-3-7-sonnet-20250219"
|
|
19
|
+
});
|
|
20
|
+
```
|
|
21
|
+
|
|
22
|
+
> **Note:** Mastra uses the OpenAI-compatible `/chat/completions` endpoint. Some provider-specific features may not be available. Check the [Merge Gateway documentation](https://docs.merge.dev/merge-gateway) for details.
|
|
23
|
+
|
|
24
|
+
## Configuration
|
|
25
|
+
|
|
26
|
+
```bash
|
|
27
|
+
# Use gateway API key
|
|
28
|
+
MERGE_GATEWAY_API_KEY=your-gateway-key
|
|
29
|
+
|
|
30
|
+
# Or use provider API keys directly
|
|
31
|
+
OPENAI_API_KEY=sk-...
|
|
32
|
+
ANTHROPIC_API_KEY=ant-...
|
|
33
|
+
```
|
|
34
|
+
|
|
35
|
+
## Available models
|
|
36
|
+
|
|
37
|
+
| Model |
|
|
38
|
+
| ------------------------------------------------ |
|
|
39
|
+
| `anthropic/claude-3-7-sonnet-20250219` |
|
|
40
|
+
| `anthropic/claude-fable-5` |
|
|
41
|
+
| `anthropic/claude-haiku-4-5-20251001` |
|
|
42
|
+
| `anthropic/claude-opus-4-1-20250805` |
|
|
43
|
+
| `anthropic/claude-opus-4-20250514` |
|
|
44
|
+
| `anthropic/claude-opus-4-5-20251101` |
|
|
45
|
+
| `anthropic/claude-opus-4-6` |
|
|
46
|
+
| `anthropic/claude-opus-4-7` |
|
|
47
|
+
| `anthropic/claude-opus-4-8` |
|
|
48
|
+
| `anthropic/claude-opus-5` |
|
|
49
|
+
| `anthropic/claude-sonnet-4-20250514` |
|
|
50
|
+
| `anthropic/claude-sonnet-4-5-20250929` |
|
|
51
|
+
| `anthropic/claude-sonnet-4-6` |
|
|
52
|
+
| `anthropic/claude-sonnet-5` |
|
|
53
|
+
| `bytedance/dola-seed-2.0-code` |
|
|
54
|
+
| `bytedance/dola-seed-2.0-code-preview` |
|
|
55
|
+
| `bytedance/dola-seed-2.0-lite` |
|
|
56
|
+
| `bytedance/dola-seed-2.0-mini` |
|
|
57
|
+
| `bytedance/dola-seed-2.0-pro` |
|
|
58
|
+
| `cohere/command-a-03-2025` |
|
|
59
|
+
| `cohere/command-r-08-2024` |
|
|
60
|
+
| `cohere/command-r-plus-08-2024` |
|
|
61
|
+
| `cohere/command-r7b-12-2024` |
|
|
62
|
+
| `deepseek/deepseek-r1` |
|
|
63
|
+
| `deepseek/deepseek-v3` |
|
|
64
|
+
| `deepseek/deepseek-v3.1` |
|
|
65
|
+
| `deepseek/deepseek-v3.2` |
|
|
66
|
+
| `deepseek/deepseek-v4-flash` |
|
|
67
|
+
| `deepseek/deepseek-v4-flash-0731` |
|
|
68
|
+
| `deepseek/deepseek-v4-pro` |
|
|
69
|
+
| `deepseek/deepseek-v4-pro-0813` |
|
|
70
|
+
| `google/gemini-2.5-computer-use-preview-10-2025` |
|
|
71
|
+
| `google/gemini-2.5-flash` |
|
|
72
|
+
| `google/gemini-2.5-flash-image` |
|
|
73
|
+
| `google/gemini-2.5-flash-lite` |
|
|
74
|
+
| `google/gemini-2.5-pro` |
|
|
75
|
+
| `google/gemini-3-flash-preview` |
|
|
76
|
+
| `google/gemini-3-pro-image` |
|
|
77
|
+
| `google/gemini-3-pro-preview` |
|
|
78
|
+
| `google/gemini-3.1-flash-image` |
|
|
79
|
+
| `google/gemini-3.1-flash-lite` |
|
|
80
|
+
| `google/gemini-3.1-flash-lite-preview` |
|
|
81
|
+
| `google/gemini-3.1-pro-preview` |
|
|
82
|
+
| `google/gemini-3.1-pro-preview-customtools` |
|
|
83
|
+
| `google/gemini-3.5-flash` |
|
|
84
|
+
| `google/gemini-3.5-flash-lite` |
|
|
85
|
+
| `google/gemini-3.6-flash` |
|
|
86
|
+
| `google/gemini-3.7-flash` |
|
|
87
|
+
| `google/gemini-embedding-001` |
|
|
88
|
+
| `google/gemini-flash-latest` |
|
|
89
|
+
| `google/gemini-flash-lite-latest` |
|
|
90
|
+
| `google/gemma-4-26b-a4b-it` |
|
|
91
|
+
| `google/gemma-4-31b-it` |
|
|
92
|
+
| `meta/llama-3.1-8b-instruct` |
|
|
93
|
+
| `meta/llama-3.3-70b-instruct` |
|
|
94
|
+
| `meta/muse-spark-1.1` |
|
|
95
|
+
| `meta/muse-spark-1.2` |
|
|
96
|
+
| `minimax/minimax-m2` |
|
|
97
|
+
| `minimax/minimax-m2.1` |
|
|
98
|
+
| `minimax/minimax-m2.5` |
|
|
99
|
+
| `minimax/minimax-m2.5-highspeed` |
|
|
100
|
+
| `minimax/minimax-m2.7` |
|
|
101
|
+
| `minimax/minimax-m2.7-highspeed` |
|
|
102
|
+
| `minimax/minimax-m3` |
|
|
103
|
+
| `mistral/codestral-latest` |
|
|
104
|
+
| `mistral/devstral-2512` |
|
|
105
|
+
| `mistral/devstral-medium-2507` |
|
|
106
|
+
| `mistral/devstral-medium-latest` |
|
|
107
|
+
| `mistral/devstral-small-2507` |
|
|
108
|
+
| `mistral/magistral-medium-latest` |
|
|
109
|
+
| `mistral/mistral-large-2411` |
|
|
110
|
+
| `mistral/mistral-large-2512` |
|
|
111
|
+
| `mistral/mistral-large-latest` |
|
|
112
|
+
| `mistral/mistral-medium-2505` |
|
|
113
|
+
| `mistral/mistral-medium-latest` |
|
|
114
|
+
| `mistral/mistral-small-latest` |
|
|
115
|
+
| `mistral/pixtral-large-latest` |
|
|
116
|
+
| `moonshot/kimi-k2.5` |
|
|
117
|
+
| `moonshot/kimi-k2.6` |
|
|
118
|
+
| `moonshot/kimi-k2.7-code` |
|
|
119
|
+
| `moonshot/kimi-k2.7-code-highspeed` |
|
|
120
|
+
| `moonshot/kimi-k3` |
|
|
121
|
+
| `moonshotai/kimi-k2-thinking` |
|
|
122
|
+
| `nvidia/nemotron-3.5-lightning-30b-a3b` |
|
|
123
|
+
| `nvidia/nemotron-nano-9b-v2` |
|
|
124
|
+
| `openai/gpt-3.5-turbo` |
|
|
125
|
+
| `openai/gpt-4` |
|
|
126
|
+
| `openai/gpt-4-turbo` |
|
|
127
|
+
| `openai/gpt-4.1` |
|
|
128
|
+
| `openai/gpt-4.1-mini` |
|
|
129
|
+
| `openai/gpt-4.1-nano` |
|
|
130
|
+
| `openai/gpt-4o` |
|
|
131
|
+
| `openai/gpt-4o-2024-05-13` |
|
|
132
|
+
| `openai/gpt-4o-2024-08-06` |
|
|
133
|
+
| `openai/gpt-4o-2024-11-20` |
|
|
134
|
+
| `openai/gpt-4o-mini` |
|
|
135
|
+
| `openai/gpt-5` |
|
|
136
|
+
| `openai/gpt-5-chat-latest` |
|
|
137
|
+
| `openai/gpt-5-mini` |
|
|
138
|
+
| `openai/gpt-5-nano` |
|
|
139
|
+
| `openai/gpt-5.1` |
|
|
140
|
+
| `openai/gpt-5.1-chat-latest` |
|
|
141
|
+
| `openai/gpt-5.2` |
|
|
142
|
+
| `openai/gpt-5.2-chat-latest` |
|
|
143
|
+
| `openai/gpt-5.3-chat-latest` |
|
|
144
|
+
| `openai/gpt-5.4` |
|
|
145
|
+
| `openai/gpt-5.4-mini` |
|
|
146
|
+
| `openai/gpt-5.4-nano` |
|
|
147
|
+
| `openai/gpt-5.5` |
|
|
148
|
+
| `openai/gpt-5.6-luna` |
|
|
149
|
+
| `openai/gpt-5.6-sol` |
|
|
150
|
+
| `openai/gpt-5.6-terra` |
|
|
151
|
+
| `openai/gpt-oss-120b` |
|
|
152
|
+
| `openai/gpt-oss-20b` |
|
|
153
|
+
| `openai/gpt-oss-safeguard-120b` |
|
|
154
|
+
| `openai/o1` |
|
|
155
|
+
| `openai/o3` |
|
|
156
|
+
| `openai/o3-mini` |
|
|
157
|
+
| `openai/o4-mini` |
|
|
158
|
+
| `qwen/qwen-flash` |
|
|
159
|
+
| `qwen/qwen-plus` |
|
|
160
|
+
| `qwen/qwen3-235b-a22b` |
|
|
161
|
+
| `qwen/qwen3-235b-a22b-instruct-2507` |
|
|
162
|
+
| `qwen/qwen3-30b-a3b` |
|
|
163
|
+
| `qwen/qwen3-32b` |
|
|
164
|
+
| `qwen/qwen3-coder-480b-a35b-instruct` |
|
|
165
|
+
| `qwen/qwen3-coder-flash` |
|
|
166
|
+
| `qwen/qwen3-coder-next` |
|
|
167
|
+
| `qwen/qwen3-coder-plus` |
|
|
168
|
+
| `qwen/qwen3-max` |
|
|
169
|
+
| `qwen/qwen3-next-80b-a3b-instruct` |
|
|
170
|
+
| `qwen/qwen3-next-80b-a3b-thinking` |
|
|
171
|
+
| `qwen/qwen3-vl-235b-a22b-instruct` |
|
|
172
|
+
| `qwen/qwen3-vl-235b-a22b-thinking` |
|
|
173
|
+
| `qwen/qwen3-vl-plus` |
|
|
174
|
+
| `qwen/qwen3.5-122b-a10b` |
|
|
175
|
+
| `qwen/qwen3.5-27b` |
|
|
176
|
+
| `qwen/qwen3.5-35b-a3b` |
|
|
177
|
+
| `qwen/qwen3.5-397b-a17b` |
|
|
178
|
+
| `qwen/qwen3.5-9b` |
|
|
179
|
+
| `qwen/qwen3.5-flash` |
|
|
180
|
+
| `qwen/qwen3.5-plus` |
|
|
181
|
+
| `qwen/qwen3.6-27b` |
|
|
182
|
+
| `qwen/qwen3.6-35b-a3b` |
|
|
183
|
+
| `qwen/qwen3.6-flash` |
|
|
184
|
+
| `qwen/qwen3.6-max-preview` |
|
|
185
|
+
| `qwen/qwen3.6-plus` |
|
|
186
|
+
| `qwen/qwen3.7-max` |
|
|
187
|
+
| `qwen/qwen3.7-plus` |
|
|
188
|
+
| `qwen/qwen3.8-2.4t-a95b` |
|
|
189
|
+
| `qwen/qwen3.8-max` |
|
|
190
|
+
| `sakana/fugu-ultra` |
|
|
191
|
+
| `sakana/sakana-namazu` |
|
|
192
|
+
| `thinkingmachines/inkling` |
|
|
193
|
+
| `writer/palmyra-x4` |
|
|
194
|
+
| `writer/palmyra-x5` |
|
|
195
|
+
| `xai/grok-4.20-0309-non-reasoning` |
|
|
196
|
+
| `xai/grok-4.20-0309-reasoning` |
|
|
197
|
+
| `xai/grok-4.3` |
|
|
198
|
+
| `xai/grok-4.5` |
|
|
199
|
+
| `xai/grok-4.6` |
|
|
200
|
+
| `xai/grok-build-0.1` |
|
|
201
|
+
| `zai/glm-4.5` |
|
|
202
|
+
| `zai/glm-4.5-air` |
|
|
203
|
+
| `zai/glm-4.5v` |
|
|
204
|
+
| `zai/glm-4.6` |
|
|
205
|
+
| `zai/glm-4.7` |
|
|
206
|
+
| `zai/glm-4.7-flash` |
|
|
207
|
+
| `zai/glm-4.7-flashx` |
|
|
208
|
+
| `zai/glm-5` |
|
|
209
|
+
| `zai/glm-5-turbo` |
|
|
210
|
+
| `zai/glm-5.1` |
|
|
211
|
+
| `zai/glm-5.2` |
|
|
212
|
+
| `zai/glm-5.3` |
|
|
@@ -2,7 +2,7 @@
|
|
|
2
2
|
|
|
3
3
|
# OpenRouter
|
|
4
4
|
|
|
5
|
-
OpenRouter aggregates models from multiple providers with enhanced features like rate limiting and failover. Access
|
|
5
|
+
OpenRouter aggregates models from multiple providers with enhanced features like rate limiting and failover. Access 353 models through Mastra's model router.
|
|
6
6
|
|
|
7
7
|
Learn more in the [OpenRouter documentation](https://openrouter.ai/models).
|
|
8
8
|
|
|
@@ -149,6 +149,7 @@ ANTHROPIC_API_KEY=ant-...
|
|
|
149
149
|
| `kwaipilot/kat-coder-air-v2.5` |
|
|
150
150
|
| `kwaipilot/kat-coder-pro-v2` |
|
|
151
151
|
| `kwaipilot/kat-coder-pro-v2.5` |
|
|
152
|
+
| `liquid/lfm-2.5-2.6b:free` |
|
|
152
153
|
| `mancer/weaver` |
|
|
153
154
|
| `meituan/longcat-2.0` |
|
|
154
155
|
| `meta-llama/llama-3.1-70b-instruct` |
|
|
@@ -385,4 +386,6 @@ ANTHROPIC_API_KEY=ant-...
|
|
|
385
386
|
| `z-ai/glm-5-turbo` |
|
|
386
387
|
| `z-ai/glm-5.1` |
|
|
387
388
|
| `z-ai/glm-5.2` |
|
|
389
|
+
| `z-ai/glm-5.2:free` |
|
|
390
|
+
| `z-ai/glm-5.3` |
|
|
388
391
|
| `z-ai/glm-5v-turbo` |
|
|
@@ -2,7 +2,7 @@
|
|
|
2
2
|
|
|
3
3
|
# Vercel
|
|
4
4
|
|
|
5
|
-
Vercel aggregates models from multiple providers with enhanced features like rate limiting and failover. Access
|
|
5
|
+
Vercel aggregates models from multiple providers with enhanced features like rate limiting and failover. Access 348 models through Mastra's model router.
|
|
6
6
|
|
|
7
7
|
Learn more in the [Vercel documentation](https://ai-sdk.dev/providers/ai-sdk-providers).
|
|
8
8
|
|
|
@@ -382,4 +382,5 @@ ANTHROPIC_API_KEY=ant-...
|
|
|
382
382
|
| `zai/glm-5.1` |
|
|
383
383
|
| `zai/glm-5.2` |
|
|
384
384
|
| `zai/glm-5.2-fast` |
|
|
385
|
+
| `zai/glm-5.3` |
|
|
385
386
|
| `zai/glm-5v-turbo` |
|
package/.docs/models/gateways.md
CHANGED
|
@@ -12,6 +12,7 @@ Create custom gateways for private LLM deployments or specialized provider integ
|
|
|
12
12
|
|
|
13
13
|
- [Azure OpenAI](https://mastra.ai/models/gateways/azure-openai)
|
|
14
14
|
- [Mastra](https://mastra.ai/models/gateways/mastra)
|
|
15
|
+
- [Merge Gateway](https://mastra.ai/models/gateways/merge-gateway)
|
|
15
16
|
- [Neon](https://mastra.ai/models/gateways/neon)
|
|
16
17
|
- [Netlify](https://mastra.ai/models/gateways/netlify)
|
|
17
18
|
- [OpenRouter](https://mastra.ai/models/gateways/openrouter)
|
package/.docs/models/index.md
CHANGED
|
@@ -2,7 +2,7 @@
|
|
|
2
2
|
|
|
3
3
|
# Model Providers
|
|
4
4
|
|
|
5
|
-
Mastra provides a unified interface for working with LLMs across multiple providers, giving you access to
|
|
5
|
+
Mastra provides a unified interface for working with LLMs across multiple providers, giving you access to 6106 models from 178 providers through a single API.
|
|
6
6
|
|
|
7
7
|
## Features
|
|
8
8
|
|
|
@@ -36,14 +36,14 @@ for await (const chunk of stream) {
|
|
|
36
36
|
|
|
37
37
|
| Model | Context | Tools | Reasoning | Image | Audio | Video | Input $/1M | Output $/1M |
|
|
38
38
|
| ----------------------------------------- | ------- | ----- | --------- | ----- | ----- | ----- | ---------- | ----------- |
|
|
39
|
-
| `ambient/ambient/large` | 203K | | | | | | $
|
|
39
|
+
| `ambient/ambient/large` | 203K | | | | | | $0.60 | $2 |
|
|
40
40
|
| `ambient/deepseek/deepseek-v4-flash` | 1.0M | | | | | | $0.14 | $0.28 |
|
|
41
41
|
| `ambient/deepseek/deepseek-v4-flash-0731` | 1.0M | | | | | | $0.14 | $0.28 |
|
|
42
42
|
| `ambient/moonshotai/kimi-k2.6` | 262K | | | | | | $0.95 | $4 |
|
|
43
43
|
| `ambient/moonshotai/kimi-k2.7-code` | 262K | | | | | | $0.69 | $3 |
|
|
44
44
|
| `ambient/stepfun/step-3.7-flash` | 262K | | | | | | $0.19 | $1 |
|
|
45
45
|
| `ambient/xiaomi/mimo-v2.5` | 1.0M | | | | | | $0.40 | $2 |
|
|
46
|
-
| `ambient/z-ai/glm-5.2` | 203K | | | | | | $
|
|
46
|
+
| `ambient/z-ai/glm-5.2` | 203K | | | | | | $0.60 | $2 |
|
|
47
47
|
| `ambient/zai-org/GLM-5.1-FP8` | 203K | | | | | | $1 | $4 |
|
|
48
48
|
| `ambient/zai-org/GLM-5.2-FP8` | 203K | | | | | | $1 | $4 |
|
|
49
49
|
|
|
@@ -46,7 +46,7 @@ for await (const chunk of stream) {
|
|
|
46
46
|
| `chutes/Qwen/Qwen3-32B-TEE` | 41K | | | | | | $0.10 | $0.42 |
|
|
47
47
|
| `chutes/Qwen/Qwen3.5-397B-A17B-TEE` | 262K | | | | | | $0.45 | $3 |
|
|
48
48
|
| `chutes/Qwen/Qwen3.6-27B-TEE` | 262K | | | | | | $0.30 | $2 |
|
|
49
|
-
| `chutes/Qwen/Qwen3.8-27B-TEE` | 262K | | | | | | $0.
|
|
49
|
+
| `chutes/Qwen/Qwen3.8-27B-TEE` | 262K | | | | | | $0.45 | $3 |
|
|
50
50
|
| `chutes/unsloth/Mistral-Nemo-Instruct-2407-TEE` | 131K | | | | | | $0.02 | $0.10 |
|
|
51
51
|
| `chutes/zai-org/GLM-5.1-TEE` | 203K | | | | | | $0.98 | $3 |
|
|
52
52
|
| `chutes/zai-org/GLM-5.2-TEE` | 1.0M | | | | | | $1 | $4 |
|
|
@@ -100,8 +100,8 @@ for await (const chunk of stream) {
|
|
|
100
100
|
| `edenai/deepinfra/zai-org/GLM-4.7-Flash` | 203K | | | | | | $0.06 | $0.40 |
|
|
101
101
|
| `edenai/deepseek/deepseek-chat` | 131K | | | | | | $0.28 | $0.42 |
|
|
102
102
|
| `edenai/deepseek/deepseek-reasoner` | 131K | | | | | | $0.28 | $0.42 |
|
|
103
|
-
| `edenai/deepseek/deepseek-v4-flash` | 1.0M | | | | | | $0.
|
|
104
|
-
| `edenai/deepseek/deepseek-v4-pro` | 1.0M | | | | | | $
|
|
103
|
+
| `edenai/deepseek/deepseek-v4-flash` | 1.0M | | | | | | $0.44 | $1 |
|
|
104
|
+
| `edenai/deepseek/deepseek-v4-pro` | 1.0M | | | | | | $1 | $4 |
|
|
105
105
|
| `edenai/fireworks_ai/accounts/fireworks/models/deepseek-v4-flash-0731` | 1.0M | | | | | | $0.14 | $0.28 |
|
|
106
106
|
| `edenai/fireworks_ai/accounts/fireworks/models/deepseek-v4-pro-0813` | 1.0M | | | | | | $1 | $4 |
|
|
107
107
|
| `edenai/fireworks_ai/accounts/fireworks/models/gpt-oss-120b` | 131K | | | | | | $0.15 | $0.60 |
|
|
@@ -158,7 +158,7 @@ for await (const chunk of stream) {
|
|
|
158
158
|
| `edenai/moonshot/kimi-k3` | 1.0M | | | | | | $3 | $15 |
|
|
159
159
|
| `edenai/nebius/meta-llama/Llama-3.3-70B-Instruct` | 131K | | | | | | $0.13 | $0.40 |
|
|
160
160
|
| `edenai/nebius/nvidia/nemotron-3-super-120b-a12b` | 8K | | | | | | $0.30 | $0.90 |
|
|
161
|
-
| `edenai/nebius/nvidia/Nemotron-3-Ultra-550b-a55b` |
|
|
161
|
+
| `edenai/nebius/nvidia/Nemotron-3-Ultra-550b-a55b` | 1.0M | | | | | | $1 | $3 |
|
|
162
162
|
| `edenai/nebius/openai/gpt-oss-120b` | 131K | | | | | | $0.15 | $0.60 |
|
|
163
163
|
| `edenai/openai/gpt-3.5-turbo` | 16K | | | | | | $0.50 | $2 |
|
|
164
164
|
| `edenai/openai/gpt-4` | 8K | | | | | | $30 | $60 |
|
|
@@ -203,7 +203,7 @@ for await (const chunk of stream) {
|
|
|
203
203
|
| `edenai/perplexityai/sonar-pro` | 200K | | | | | | $3 | $15 |
|
|
204
204
|
| `edenai/perplexityai/sonar-reasoning-pro` | 128K | | | | | | $2 | $8 |
|
|
205
205
|
| `edenai/qwen/deepseek-v4-flash-0731` | 1.0M | | | | | | $0.20 | $0.40 |
|
|
206
|
-
| `edenai/qwen/deepseek-v4-pro-0813` | 1.0M | | | | | | $
|
|
206
|
+
| `edenai/qwen/deepseek-v4-pro-0813` | 1.0M | | | | | | $0.66 | $2 |
|
|
207
207
|
| `edenai/qwen/qwen-max` | 33K | | | | | | $2 | $6 |
|
|
208
208
|
| `edenai/qwen/qwen-vl-max` | 131K | | | | | | $0.80 | $3 |
|
|
209
209
|
| `edenai/qwen/qwen-vl-plus` | 131K | | | | | | $0.21 | $0.63 |
|
|
@@ -222,7 +222,7 @@ for await (const chunk of stream) {
|
|
|
222
222
|
| `edenai/qwen/qwen3.8-max` | 1.0M | | | | | | $2 | $6 |
|
|
223
223
|
| `edenai/qwen/qwq-plus` | 131K | | | | | | $0.80 | $2 |
|
|
224
224
|
| `edenai/scaleway/deepseek-v4-flash-0731` | 256K | | | | | | $0.46 | $0.93 |
|
|
225
|
-
| `edenai/scaleway/gpt-oss-120b` | 128K | | | | | | $0.17 | $0.
|
|
225
|
+
| `edenai/scaleway/gpt-oss-120b` | 128K | | | | | | $0.17 | $0.69 |
|
|
226
226
|
| `edenai/scaleway/llama-3.3-70b-instruct` | 128K | | | | | | $1 | $1 |
|
|
227
227
|
| `edenai/tensorx/deepseek/deepseek-v4-flash-0731` | 1.0M | | | | | | $0.25 | $0.30 |
|
|
228
228
|
| `edenai/tensorx/moonshotai/kimi-k2.5` | 262K | | | | | | $0.50 | $3 |
|
|
@@ -2,7 +2,7 @@
|
|
|
2
2
|
|
|
3
3
|
# Hetzner
|
|
4
4
|
|
|
5
|
-
Access
|
|
5
|
+
Access 2 Hetzner models through Mastra's model router. Authentication is handled automatically using the `HETZNER_API_KEY` environment variable.
|
|
6
6
|
|
|
7
7
|
Learn more in the [Hetzner documentation](https://experiments.hetzner.com).
|
|
8
8
|
|
|
@@ -17,7 +17,7 @@ const agent = new Agent({
|
|
|
17
17
|
id: "my-agent",
|
|
18
18
|
name: "My Agent",
|
|
19
19
|
instructions: "You are a helpful assistant",
|
|
20
|
-
model: "hetzner/
|
|
20
|
+
model: "hetzner/Qwen/Qwen3.6-35B-A3B-FP8"
|
|
21
21
|
});
|
|
22
22
|
|
|
23
23
|
// Generate a response
|
|
@@ -36,10 +36,8 @@ for await (const chunk of stream) {
|
|
|
36
36
|
|
|
37
37
|
| Model | Context | Tools | Reasoning | Image | Audio | Video | Input $/1M | Output $/1M |
|
|
38
38
|
| ---------------------------------- | ------- | ----- | --------- | ----- | ----- | ----- | ---------- | ----------- |
|
|
39
|
-
| `hetzner/DeepSeek-V4-Flash-0731` | 512K | | | | | | — | — |
|
|
40
|
-
| `hetzner/GLM-5.2-NVFP4` | 512K | | | | | | — | — |
|
|
41
|
-
| `hetzner/Kimi-K2.7-Code` | 262K | | | | | | — | — |
|
|
42
39
|
| `hetzner/Qwen/Qwen3.6-35B-A3B-FP8` | 262K | | | | | | — | — |
|
|
40
|
+
| `hetzner/Qwen3.8-27B` | 262K | | | | | | — | — |
|
|
43
41
|
|
|
44
42
|
## Advanced configuration
|
|
45
43
|
|
|
@@ -51,7 +49,7 @@ const agent = new Agent({
|
|
|
51
49
|
name: "custom-agent",
|
|
52
50
|
model: {
|
|
53
51
|
url: "https://inference.hetzner.com/api/v1",
|
|
54
|
-
id: "hetzner/
|
|
52
|
+
id: "hetzner/Qwen/Qwen3.6-35B-A3B-FP8",
|
|
55
53
|
apiKey: process.env.HETZNER_API_KEY,
|
|
56
54
|
headers: {
|
|
57
55
|
"X-Custom-Header": "value"
|
|
@@ -69,8 +67,8 @@ const agent = new Agent({
|
|
|
69
67
|
model: ({ requestContext }) => {
|
|
70
68
|
const useAdvanced = requestContext.task === "complex";
|
|
71
69
|
return useAdvanced
|
|
72
|
-
? "hetzner/
|
|
73
|
-
: "hetzner/
|
|
70
|
+
? "hetzner/Qwen3.8-27B"
|
|
71
|
+
: "hetzner/Qwen/Qwen3.6-35B-A3B-FP8";
|
|
74
72
|
}
|
|
75
73
|
});
|
|
76
74
|
```
|
|
@@ -41,17 +41,17 @@ for await (const chunk of stream) {
|
|
|
41
41
|
| `hyper/deepseek-v4-pro` | 1.0M | | | | | | $2 | $5 |
|
|
42
42
|
| `hyper/deepseek-v4-pro-0813` | 1.0M | | | | | | $1 | $4 |
|
|
43
43
|
| `hyper/gemma-4-26b-a4b-it` | 256K | | | | | | $0.12 | $0.42 |
|
|
44
|
-
| `hyper/glm-5` | 203K | | | | | | $0.
|
|
45
|
-
| `hyper/glm-5.1` | 203K | | | | | | $
|
|
44
|
+
| `hyper/glm-5` | 203K | | | | | | $0.84 | $3 |
|
|
45
|
+
| `hyper/glm-5.1` | 203K | | | | | | $1 | $4 |
|
|
46
46
|
| `hyper/glm-5.2` | 1.0M | | | | | | $1 | $4 |
|
|
47
47
|
| `hyper/gpt-oss-120b` | 131K | | | | | | $0.16 | $0.65 |
|
|
48
|
-
| `hyper/kimi-k2.5` | 262K | | | | | | $0.
|
|
48
|
+
| `hyper/kimi-k2.5` | 262K | | | | | | $0.57 | $3 |
|
|
49
49
|
| `hyper/kimi-k2.6` | 262K | | | | | | $0.95 | $4 |
|
|
50
50
|
| `hyper/kimi-k2.7-code` | 256K | | | | | | $0.95 | $4 |
|
|
51
51
|
| `hyper/kimi-k3` | 1.0M | | | | | | $3 | $15 |
|
|
52
|
-
| `hyper/llama-3.3-70b-instruct` | 128K | | | | | | $0.
|
|
52
|
+
| `hyper/llama-3.3-70b-instruct` | 128K | | | | | | $0.64 | $0.77 |
|
|
53
53
|
| `hyper/llama-4-maverick-17b-128e-instruct-fp8` | 430K | | | | | | $0.27 | $0.90 |
|
|
54
|
-
| `hyper/minimax-m2.7` | 262K | | | | | | $0.
|
|
54
|
+
| `hyper/minimax-m2.7` | 262K | | | | | | $0.47 | $2 |
|
|
55
55
|
| `hyper/minimax-m3` | 512K | | | | | | $0.33 | $1 |
|
|
56
56
|
| `hyper/qwen3-coder-480b-a35b-instruct-int4-mixed-ar` | 106K | | | | | | $0.45 | $2 |
|
|
57
57
|
| `hyper/qwen3-next-80b-a3b-instruct` | 262K | | | | | | $0.12 | $1 |
|
|
@@ -2,7 +2,7 @@
|
|
|
2
2
|
|
|
3
3
|
# Kilo Gateway
|
|
4
4
|
|
|
5
|
-
Access
|
|
5
|
+
Access 360 Kilo Gateway models through Mastra's model router. Authentication is handled automatically using the `KILO_API_KEY` environment variable.
|
|
6
6
|
|
|
7
7
|
Learn more in the [Kilo Gateway documentation](https://kilo.ai).
|
|
8
8
|
|
|
@@ -40,7 +40,7 @@ for await (const chunk of stream) {
|
|
|
40
40
|
| `kilo/~anthropic/claude-haiku-latest` | 200K | | | | | | $1 | $5 |
|
|
41
41
|
| `kilo/~anthropic/claude-opus-latest` | 1.0M | | | | | | $5 | $25 |
|
|
42
42
|
| `kilo/~anthropic/claude-sonnet-latest` | 1.0M | | | | | | $2 | $10 |
|
|
43
|
-
| `kilo/~deepseek/deepseek-v4-flash-latest` |
|
|
43
|
+
| `kilo/~deepseek/deepseek-v4-flash-latest` | 262K | | | | | | $0.08 | $0.15 |
|
|
44
44
|
| `kilo/~google/gemini-flash-latest` | 1.0M | | | | | | $0.38 | $2 |
|
|
45
45
|
| `kilo/~google/gemini-pro-latest` | 1.0M | | | | | | $2 | $12 |
|
|
46
46
|
| `kilo/~moonshotai/kimi-latest` | 975K | | | | | | $3 | $13 |
|
|
@@ -128,7 +128,7 @@ for await (const chunk of stream) {
|
|
|
128
128
|
| `kilo/google/gemini-3.5-flash` | 1.0M | | | | | | $2 | $9 |
|
|
129
129
|
| `kilo/google/gemini-3.5-flash-lite` | 1.0M | | | | | | $0.30 | $3 |
|
|
130
130
|
| `kilo/google/gemini-3.6-flash` | 1.0M | | | | | | $0.75 | $4 |
|
|
131
|
-
| `kilo/google/gemini-3.7-flash` | 1.0M | | | | | | $
|
|
131
|
+
| `kilo/google/gemini-3.7-flash` | 1.0M | | | | | | $2 | $8 |
|
|
132
132
|
| `kilo/google/gemma-2-27b-it` | 8K | | | | | | $0.65 | $0.65 |
|
|
133
133
|
| `kilo/google/gemma-3-12b-it` | 131K | | | | | | $0.05 | $0.15 |
|
|
134
134
|
| `kilo/google/gemma-3-27b-it` | 131K | | | | | | $0.08 | $0.16 |
|
|
@@ -154,6 +154,7 @@ for await (const chunk of stream) {
|
|
|
154
154
|
| `kilo/kwaipilot/kat-coder-air-v2.5` | 256K | | | | | | $0.15 | $0.60 |
|
|
155
155
|
| `kilo/kwaipilot/kat-coder-pro-v2` | 256K | | | | | | $0.30 | $1 |
|
|
156
156
|
| `kilo/kwaipilot/kat-coder-pro-v2.5` | 256K | | | | | | $0.74 | $3 |
|
|
157
|
+
| `kilo/liquid/lfm-2.5-2.6b:free` | 128K | | | | | | — | — |
|
|
157
158
|
| `kilo/mancer/weaver` | 8K | | | | | | $0.50 | $0.75 |
|
|
158
159
|
| `kilo/meituan/longcat-2.0` | 1.0M | | | | | | $0.75 | $3 |
|
|
159
160
|
| `kilo/meta-llama/llama-3.1-70b-instruct` | 131K | | | | | | $0.40 | $0.40 |
|
|
@@ -393,6 +394,7 @@ for await (const chunk of stream) {
|
|
|
393
394
|
| `kilo/z-ai/glm-5-turbo` | 203K | | | | | | $1 | $4 |
|
|
394
395
|
| `kilo/z-ai/glm-5.1` | 200K | | | | | | $1 | $4 |
|
|
395
396
|
| `kilo/z-ai/glm-5.2` | 1.0M | | | | | | $1 | $4 |
|
|
397
|
+
| `kilo/z-ai/glm-5.3` | 1.0M | | | | | | $1 | $4 |
|
|
396
398
|
| `kilo/z-ai/glm-5v-turbo` | 203K | | | | | | $1 | $4 |
|
|
397
399
|
|
|
398
400
|
## Advanced configuration
|
|
@@ -55,7 +55,7 @@ for await (const chunk of stream) {
|
|
|
55
55
|
| `llmgateway/cosmos3-super-reasoner` | 262K | | | | | | $0.10 | $0.30 |
|
|
56
56
|
| `llmgateway/custom` | 128K | | | | | | — | — |
|
|
57
57
|
| `llmgateway/deepseek-v3.2` | 164K | | | | | | $0.26 | $0.38 |
|
|
58
|
-
| `llmgateway/deepseek-v4-flash` | 1.1M | | | | | | $0.
|
|
58
|
+
| `llmgateway/deepseek-v4-flash` | 1.1M | | | | | | $0.05 | $0.09 |
|
|
59
59
|
| `llmgateway/deepseek-v4-pro` | 1.1M | | | | | | $0.43 | $0.87 |
|
|
60
60
|
| `llmgateway/ernie-4.5-vl-424b-a47b` | 123K | | | | | | $0.42 | $1 |
|
|
61
61
|
| `llmgateway/fugu-ultra` | 1.0M | | | | | | $5 | $30 |
|
|
@@ -176,7 +176,7 @@ for await (const chunk of stream) {
|
|
|
176
176
|
| `llmgateway/nemotron-3-nano-30b` | 262K | | | | | | $0.06 | $0.24 |
|
|
177
177
|
| `llmgateway/nemotron-3-nano-omni` | 262K | | | | | | $0.06 | $0.24 |
|
|
178
178
|
| `llmgateway/nemotron-3-super-120b` | 262K | | | | | | $0.30 | $0.90 |
|
|
179
|
-
| `llmgateway/nemotron-3-ultra-550b` | 1.0M | | | | | | $0.50 | $
|
|
179
|
+
| `llmgateway/nemotron-3-ultra-550b` | 1.0M | | | | | | $0.50 | $2 |
|
|
180
180
|
| `llmgateway/o1` | 200K | | | | | | $15 | $60 |
|
|
181
181
|
| `llmgateway/o3` | 200K | | | | | | $2 | $8 |
|
|
182
182
|
| `llmgateway/o3-mini` | 200K | | | | | | $1 | $4 |
|
|
@@ -194,7 +194,7 @@ for await (const chunk of stream) {
|
|
|
194
194
|
| `llmgateway/qwen3-30b-a3b-instruct-2507` | 262K | | | | | | $0.10 | $0.30 |
|
|
195
195
|
| `llmgateway/qwen3-32b` | 41K | | | | | | $0.10 | $0.30 |
|
|
196
196
|
| `llmgateway/qwen3-coder-30b-a3b-instruct` | 262K | | | | | | $0.07 | $0.27 |
|
|
197
|
-
| `llmgateway/qwen3-coder-480b-a35b-instruct` | 262K | | | | | | $0.
|
|
197
|
+
| `llmgateway/qwen3-coder-480b-a35b-instruct` | 262K | | | | | | $0.38 | $2 |
|
|
198
198
|
| `llmgateway/qwen3-coder-flash` | 1.0M | | | | | | $0.30 | $2 |
|
|
199
199
|
| `llmgateway/qwen3-coder-next` | 262K | | | | | | $0.11 | $0.68 |
|
|
200
200
|
| `llmgateway/qwen3-coder-plus` | 1.0M | | | | | | $6 | $60 |
|
|
@@ -0,0 +1,76 @@
|
|
|
1
|
+
> Discover all available pages from the documentation index: https://mastra.ai/llms.txt
|
|
2
|
+
|
|
3
|
+
# SCX.ai
|
|
4
|
+
|
|
5
|
+
Access 4 SCX.ai models through Mastra's model router. Authentication is handled automatically using the `SCX_API_KEY` environment variable.
|
|
6
|
+
|
|
7
|
+
Learn more in the [SCX.ai documentation](https://platform.scx.ai/docs).
|
|
8
|
+
|
|
9
|
+
```bash
|
|
10
|
+
SCX_API_KEY=your-api-key
|
|
11
|
+
```
|
|
12
|
+
|
|
13
|
+
```typescript
|
|
14
|
+
import { Agent } from "@mastra/core/agent";
|
|
15
|
+
|
|
16
|
+
const agent = new Agent({
|
|
17
|
+
id: "my-agent",
|
|
18
|
+
name: "My Agent",
|
|
19
|
+
instructions: "You are a helpful assistant",
|
|
20
|
+
model: "scx-ai/GLM-5.2"
|
|
21
|
+
});
|
|
22
|
+
|
|
23
|
+
// Generate a response
|
|
24
|
+
const response = await agent.generate("Hello!");
|
|
25
|
+
|
|
26
|
+
// Stream a response
|
|
27
|
+
const stream = await agent.stream("Tell me a story");
|
|
28
|
+
for await (const chunk of stream) {
|
|
29
|
+
console.log(chunk);
|
|
30
|
+
}
|
|
31
|
+
```
|
|
32
|
+
|
|
33
|
+
> **Note:** Mastra uses the OpenAI-compatible `/chat/completions` endpoint. Some provider-specific features may not be available. Check the [SCX.ai documentation](https://platform.scx.ai/docs) for details.
|
|
34
|
+
|
|
35
|
+
## Models
|
|
36
|
+
|
|
37
|
+
| Model | Context | Tools | Reasoning | Image | Audio | Video | Input $/1M | Output $/1M |
|
|
38
|
+
| --------------------- | ------- | ----- | --------- | ----- | ----- | ----- | ---------- | ----------- |
|
|
39
|
+
| `scx-ai/GLM-5.2` | 1.0M | | | | | | $0.55 | $2 |
|
|
40
|
+
| `scx-ai/gpt-oss-120b` | 131K | | | | | | $0.17 | $0.55 |
|
|
41
|
+
| `scx-ai/MiniMax-M2.7` | 197K | | | | | | $0.48 | $2 |
|
|
42
|
+
| `scx-ai/Qwen3.8-Max` | 1.0M | | | | | | $2 | $5 |
|
|
43
|
+
|
|
44
|
+
## Advanced configuration
|
|
45
|
+
|
|
46
|
+
### Custom headers
|
|
47
|
+
|
|
48
|
+
```typescript
|
|
49
|
+
const agent = new Agent({
|
|
50
|
+
id: "custom-agent",
|
|
51
|
+
name: "custom-agent",
|
|
52
|
+
model: {
|
|
53
|
+
url: "https://api.scx.ai/v1",
|
|
54
|
+
id: "scx-ai/GLM-5.2",
|
|
55
|
+
apiKey: process.env.SCX_API_KEY,
|
|
56
|
+
headers: {
|
|
57
|
+
"X-Custom-Header": "value"
|
|
58
|
+
}
|
|
59
|
+
}
|
|
60
|
+
});
|
|
61
|
+
```
|
|
62
|
+
|
|
63
|
+
### Dynamic model selection
|
|
64
|
+
|
|
65
|
+
```typescript
|
|
66
|
+
const agent = new Agent({
|
|
67
|
+
id: "dynamic-agent",
|
|
68
|
+
name: "Dynamic Agent",
|
|
69
|
+
model: ({ requestContext }) => {
|
|
70
|
+
const useAdvanced = requestContext.task === "complex";
|
|
71
|
+
return useAdvanced
|
|
72
|
+
? "scx-ai/gpt-oss-120b"
|
|
73
|
+
: "scx-ai/GLM-5.2";
|
|
74
|
+
}
|
|
75
|
+
});
|
|
76
|
+
```
|
|
@@ -2,7 +2,7 @@
|
|
|
2
2
|
|
|
3
3
|
# Weights & Biases
|
|
4
4
|
|
|
5
|
-
Access
|
|
5
|
+
Access 29 Weights & Biases models through Mastra's model router. Authentication is handled automatically using the `WANDB_API_KEY` environment variable.
|
|
6
6
|
|
|
7
7
|
Learn more in the [Weights & Biases documentation](https://docs.wandb.ai).
|
|
8
8
|
|
|
@@ -62,6 +62,7 @@ for await (const chunk of stream) {
|
|
|
62
62
|
| `wandb/Qwen/Qwen3.5-35B-A3B` | 262K | | | | | | $0.25 | $1 |
|
|
63
63
|
| `wandb/Qwen/Qwen3.6-27B` | 262K | | | | | | $0.60 | $4 |
|
|
64
64
|
| `wandb/Qwen/Qwen3.6-35B-A3B` | 262K | | | | | | $0.25 | $1 |
|
|
65
|
+
| `wandb/Qwen/Qwen3.8-27B` | 262K | | | | | | $0.40 | $3 |
|
|
65
66
|
| `wandb/zai-org/GLM-5.1` | 203K | | | | | | $1 | $4 |
|
|
66
67
|
| `wandb/zai-org/GLM-5.2` | 262K | | | | | | $0.76 | $2 |
|
|
67
68
|
|
|
@@ -91,7 +91,6 @@ Direct access to individual AI model providers. Each provider offers unique mode
|
|
|
91
91
|
- [LucidQuery](https://mastra.ai/models/providers/lucidquery)
|
|
92
92
|
- [Lynkr](https://mastra.ai/models/providers/lynkr)
|
|
93
93
|
- [Meganova](https://mastra.ai/models/providers/meganova)
|
|
94
|
-
- [Merge Gateway](https://mastra.ai/models/providers/merge-gateway)
|
|
95
94
|
- [Meta](https://mastra.ai/models/providers/meta)
|
|
96
95
|
- [MiniMax (minimax.io)](https://mastra.ai/models/providers/minimax)
|
|
97
96
|
- [MiniMax (minimaxi.com)](https://mastra.ai/models/providers/minimax-cn)
|
|
@@ -135,7 +134,7 @@ Direct access to individual AI model providers. Each provider offers unique mode
|
|
|
135
134
|
- [Sarvam AI](https://mastra.ai/models/providers/sarvam)
|
|
136
135
|
- [Scaleway](https://mastra.ai/models/providers/scaleway)
|
|
137
136
|
- [SCNet Token Plan](https://mastra.ai/models/providers/scnet-token-plan)
|
|
138
|
-
- [SCX.ai](https://mastra.ai/models/providers/scx)
|
|
137
|
+
- [SCX.ai](https://mastra.ai/models/providers/scx-ai)
|
|
139
138
|
- [SiliconFlow](https://mastra.ai/models/providers/siliconflow)
|
|
140
139
|
- [SiliconFlow (China)](https://mastra.ai/models/providers/siliconflow-cn)
|
|
141
140
|
- [Snowflake Cortex](https://mastra.ai/models/providers/snowflake-cortex)
|