@mastra/mcp-docs-server 1.3.2-alpha.6 → 1.3.2-alpha.8

This diff represents the content of publicly available package versions that have been released to one of the supported registries. The information contained in this diff is provided for informational purposes only and reflects changes between package versions as they appear in their respective public registries.
@@ -278,6 +278,8 @@ Under the default persistence policy, these `running` checkpoints are only writt
278
278
 
279
279
  Durable agent runs are excluded from the generic boot-time restart of active workflow runs. The only automatic recovery path for durable agent runs is `recovery.durableAgents: 'auto'`, which holds a recovery lease and registers thread runtimes before re-driving each run.
280
280
 
281
+ > **Warning:** Automatic and manual recovery re-run the agentic loop from the last persisted snapshot, which can reissue LLM calls (with real cost) and re-execute tool calls. A sub-agent delegation is one generated `agent-<name>` tool call, and recovery doesn't checkpoint the LLM or tool calls inside that delegation independently. If recovery replays the delegation step, the entire sub-agent run starts again and can repeat inner LLM calls and already-completed tool side effects. Make tools, including tools used by sub-agents, idempotent, or deduplicate their external side effects.
282
+
281
283
  ### Snapshot persistence
282
284
 
283
285
  The `shouldPersistSnapshot` option on `createDurableAgent()` (also accepted in the agent-level `durable` config) controls which workflow snapshots a durable agent writes. The default policy:
@@ -313,8 +315,6 @@ export const mastra = new Mastra({
313
315
 
314
316
  On startup, this discovers every registered durable agent with runs stuck in `running` status and re-drives them from the last persisted snapshot.
315
317
 
316
- > **Warning:** Recovery re-runs the agentic loop from the last snapshot, which re-issues LLM calls (real cost) and re-executes tool calls. Make sure your tools are idempotent before enabling automatic recovery.
317
-
318
318
  ### Manual recovery
319
319
 
320
320
  If you need finer control, such as gating recovery behind a leader election or running it on a schedule, call the methods directly:
@@ -4,7 +4,7 @@
4
4
 
5
5
  # ![OpenRouter logo](https://models.dev/logos/openrouter.svg)OpenRouter
6
6
 
7
- OpenRouter aggregates models from multiple providers with enhanced features like rate limiting and failover. Access 385 models through Mastra's model router.
7
+ OpenRouter aggregates models from multiple providers with enhanced features like rate limiting and failover. Access 384 models through Mastra's model router.
8
8
 
9
9
  Learn more in the [OpenRouter documentation](https://openrouter.ai/models).
10
10
 
@@ -68,7 +68,6 @@ ANTHROPIC_API_KEY=ant-...
68
68
  | `amazon/nova-premier-v1` |
69
69
  | `amazon/nova-pro-v1` |
70
70
  | `anthracite-org/magnum-v4-72b` |
71
- | `anthropic/claude-3-haiku` |
72
71
  | `anthropic/claude-fable-5` |
73
72
  | `anthropic/claude-fable-5.1` |
74
73
  | `anthropic/claude-haiku-4.5` |
@@ -4,7 +4,7 @@
4
4
 
5
5
  # Model Providers
6
6
 
7
- Mastra provides a unified interface for working with LLMs across multiple providers, giving you access to 7630 models from 210 providers through a single API.
7
+ Mastra provides a unified interface for working with LLMs across multiple providers, giving you access to 7622 models from 210 providers through a single API.
8
8
 
9
9
  ## Features
10
10
 
@@ -4,7 +4,7 @@
4
4
 
5
5
  # ![Eden AI logo](https://models.dev/logos/edenai.svg)Eden AI
6
6
 
7
- Access 285 Eden AI models through Mastra's model router. Authentication is handled automatically using the `EDENAI_API_KEY` environment variable.
7
+ Access 287 Eden AI models through Mastra's model router. Authentication is handled automatically using the `EDENAI_API_KEY` environment variable.
8
8
 
9
9
  Learn more in the [Eden AI documentation](https://docs.edenai.co).
10
10
 
@@ -89,6 +89,7 @@ for await (const chunk of stream) {
89
89
  | `edenai/cloudflare/@cf/openai/gpt-oss-20b` | 128K | | | | | | $0.20 | $0.30 |
90
90
  | `edenai/cloudflare/@cf/qwen/qwen2.5-coder-32b-instruct` | 33K | | | | | | $0.66 | $1 |
91
91
  | `edenai/cloudflare/@cf/zai-org/glm-4.7-flash` | 131K | | | | | | $0.06 | $0.40 |
92
+ | `edenai/cohere/c4ai-aya-expanse-32b` | 128K | | | | | | $0.50 | $2 |
92
93
  | `edenai/cohere/command-a-03-2025` | 288K | | | | | | $3 | $10 |
93
94
  | `edenai/cohere/command-r-08-2024` | 128K | | | | | | $0.15 | $0.60 |
94
95
  | `edenai/cohere/command-r-plus-08-2024` | 128K | | | | | | $3 | $10 |
@@ -124,6 +125,7 @@ for await (const chunk of stream) {
124
125
  | `edenai/deepinfra/stepfun-ai/Step-3.5-Flash` | 262K | | | | | | $0.09 | $0.30 |
125
126
  | `edenai/deepinfra/stepfun-ai/Step-3.7-Flash` | 262K | | | | | | $0.20 | $1 |
126
127
  | `edenai/deepinfra/tencent/Hy3` | 262K | | | | | | $0.13 | $0.53 |
128
+ | `edenai/deepinfra/tencent/Hy4-preview` | 1.0M | | | | | | $0.83 | $3 |
127
129
  | `edenai/deepinfra/thinkingmachines/Inkling` | 524K | | | | | | $0.95 | $4 |
128
130
  | `edenai/deepinfra/thinkingmachines/Inkling-Small` | 524K | | | | | | $0.45 | $1 |
129
131
  | `edenai/deepinfra/zai-org/GLM-4.7-Flash` | 203K | | | | | | $0.06 | $0.40 |
@@ -4,7 +4,7 @@
4
4
 
5
5
  # ![Fireworks AI logo](https://models.dev/logos/fireworks-ai.svg)Fireworks AI
6
6
 
7
- Access 34 Fireworks AI models through Mastra's model router. Authentication is handled automatically using the `FIREWORKS_API_KEY` environment variable.
7
+ Access 22 Fireworks AI models through Mastra's model router. Authentication is handled automatically using the `FIREWORKS_API_KEY` environment variable.
8
8
 
9
9
  Learn more in the [Fireworks AI documentation](https://fireworks.ai/docs/).
10
10
 
@@ -19,7 +19,7 @@ const agent = new Agent({
19
19
  id: "my-agent",
20
20
  name: "My Agent",
21
21
  instructions: "You are a helpful assistant",
22
- model: "fireworks-ai/accounts/fireworks/models/deepseek-v4-flash-0731"
22
+ model: "fireworks-ai/accounts/fireworks/models/deepseek-v4p1-flash"
23
23
  });
24
24
 
25
25
  // Generate a response
@@ -48,12 +48,9 @@ for await (const chunk of stream) {
48
48
  | `fireworks-ai/accounts/fireworks/models/minimax-m3` | 512K | | | | | | $0.30 | $1 |
49
49
  | `fireworks-ai/accounts/fireworks/models/nemotron-3-ultra-nvfp4` | 262K | | | | | | $0.60 | $2 |
50
50
  | `fireworks-ai/accounts/fireworks/models/nemotron-lightning-3p5-30b-a3b` | 262K | | | | | | $0.05 | $0.20 |
51
- | `fireworks-ai/accounts/fireworks/models/qwen3p7-plus` | 262K | | | | | | $0.40 | $2 |
52
51
  | `fireworks-ai/accounts/fireworks/models/qwen3p8-2p4t-a95b` | 262K | | | | | | $2 | $6 |
53
52
  | `fireworks-ai/accounts/fireworks/models/qwen3p8-max` | 262K | | | | | | $2 | $6 |
54
53
  | `fireworks-ai/accounts/fireworks/routers/deepseek-flash-latest` | 1.0M | | | | | | $0.22 | $0.66 |
55
- | `fireworks-ai/accounts/fireworks/routers/deepseek-pro-latest` | 1.0M | | | | | | $1 | $4 |
56
- | `fireworks-ai/accounts/fireworks/routers/glm-5p2-fast` | 1.0M | | | | | | $2 | $7 |
57
54
  | `fireworks-ai/accounts/fireworks/routers/glm-5p3-fast` | 1.0M | | | | | | $2 | $7 |
58
55
  | `fireworks-ai/accounts/fireworks/routers/glm-fast-latest` | 1.0M | | | | | | $2 | $7 |
59
56
  | `fireworks-ai/accounts/fireworks/routers/glm-flash-latest` | 1.0M | | | | | | $0.15 | $0.50 |
@@ -76,7 +73,7 @@ const agent = new Agent({
76
73
  name: "custom-agent",
77
74
  model: {
78
75
  url: "https://api.fireworks.ai/inference/v1/",
79
- id: "fireworks-ai/accounts/fireworks/models/deepseek-v4-flash-0731",
76
+ id: "fireworks-ai/accounts/fireworks/models/deepseek-v4p1-flash",
80
77
  apiKey: process.env.FIREWORKS_API_KEY,
81
78
  headers: {
82
79
  "X-Custom-Header": "value"
@@ -95,7 +92,7 @@ const agent = new Agent({
95
92
  const useAdvanced = requestContext.task === "complex";
96
93
  return useAdvanced
97
94
  ? "fireworks-ai/accounts/fireworks/routers/qwen-max-latest"
98
- : "fireworks-ai/accounts/fireworks/models/deepseek-v4-flash-0731";
95
+ : "fireworks-ai/accounts/fireworks/models/deepseek-v4p1-flash";
99
96
  }
100
97
  });
101
98
  ```
@@ -4,7 +4,7 @@
4
4
 
5
5
  # ![Kilo Gateway logo](https://models.dev/logos/kilo.svg)Kilo Gateway
6
6
 
7
- Access 392 Kilo Gateway models through Mastra's model router. Authentication is handled automatically using the `KILO_API_KEY` environment variable.
7
+ Access 391 Kilo Gateway models through Mastra's model router. Authentication is handled automatically using the `KILO_API_KEY` environment variable.
8
8
 
9
9
  Learn more in the [Kilo Gateway documentation](https://kilo.ai).
10
10
 
@@ -43,19 +43,19 @@ for await (const chunk of stream) {
43
43
  | `kilo/~anthropic/claude-opus-latest` | 1.0M | | | | | | $4 | $20 |
44
44
  | `kilo/~anthropic/claude-sonnet-latest` | 1.0M | | | | | | $2 | $10 |
45
45
  | `kilo/~deepseek/deepseek-flash-latest` | 1.0M | | | | | | $0.04 | $0.29 |
46
- | `kilo/~deepseek/deepseek-pro-latest` | 1.0M | | | | | | $0.25 | $4 |
46
+ | `kilo/~deepseek/deepseek-pro-latest` | 1.0M | | | | | | $0.25 | $0.74 |
47
47
  | `kilo/~deepseek/deepseek-v4-flash-latest` | 1.0M | | | | | | $0.02 | $0.32 |
48
48
  | `kilo/~google/gemini-flash-latest` | 1.0M | | | | | | $0.75 | $4 |
49
49
  | `kilo/~google/gemini-pro-latest` | 1.0M | | | | | | $2 | $12 |
50
- | `kilo/~moonshotai/kimi-latest` | 1.0M | | | | | | $1 | $11 |
50
+ | `kilo/~moonshotai/kimi-latest` | 1.0M | | | | | | $1 | $9 |
51
51
  | `kilo/~openai/gpt-astra-latest` | 1.1M | | | | | | $10 | $50 |
52
52
  | `kilo/~openai/gpt-luna-latest` | 1.1M | | | | | | $0.10 | $0.50 |
53
53
  | `kilo/~openai/gpt-mini-latest` | 400K | | | | | | $0.75 | $5 |
54
54
  | `kilo/~openai/gpt-sol-latest` | 1.1M | | | | | | $2 | $10 |
55
55
  | `kilo/~openai/gpt-terra-latest` | 1.1M | | | | | | $2 | $12 |
56
56
  | `kilo/~x-ai/grok-latest` | 500K | | | | | | $2 | $5 |
57
- | `kilo/~z-ai/glm-flash-latest` | 1.0M | | | | | | $0.04 | $0.14 |
58
- | `kilo/~z-ai/glm-latest` | 1.0M | | | | | | $0.56 | $2 |
57
+ | `kilo/~z-ai/glm-flash-latest` | 1.0M | | | | | | $0.04 | $0.50 |
58
+ | `kilo/~z-ai/glm-latest` | 1.0M | | | | | | $0.38 | $1 |
59
59
  | `kilo/aion-labs/aion-2.0` | 131K | | | | | | $0.80 | $2 |
60
60
  | `kilo/aion-labs/aion-3.0` | 131K | | | | | | $3 | $6 |
61
61
  | `kilo/aion-labs/aion-3.0-mini` | 131K | | | | | | $0.70 | $1 |
@@ -68,7 +68,6 @@ for await (const chunk of stream) {
68
68
  | `kilo/amazon/nova-premier-v1` | 1.0M | | | | | | $3 | $13 |
69
69
  | `kilo/amazon/nova-pro-v1` | 300K | | | | | | $0.80 | $3 |
70
70
  | `kilo/anthracite-org/magnum-v4-72b` | 33K | | | | | | $3 | $5 |
71
- | `kilo/anthropic/claude-3-haiku` | 200K | | | | | | $0.25 | $1 |
72
71
  | `kilo/anthropic/claude-fable-5` | 1.0M | | | | | | $10 | $50 |
73
72
  | `kilo/anthropic/claude-fable-5.1` | 1.0M | | | | | | $10 | $50 |
74
73
  | `kilo/anthropic/claude-haiku-4.5` | 200K | | | | | | $1 | $5 |
@@ -141,7 +140,7 @@ for await (const chunk of stream) {
141
140
  | `kilo/google/gemma-3-27b-it` | 131K | | | | | | $0.08 | $0.16 |
142
141
  | `kilo/google/gemma-3-4b-it` | 131K | | | | | | $0.05 | $0.10 |
143
142
  | `kilo/google/gemma-4-26b-a4b-it` | 262K | | | | | | $0.04 | $0.22 |
144
- | `kilo/google/gemma-4-31b-it` | 262K | | | | | | $0.09 | $0.34 |
143
+ | `kilo/google/gemma-4-31b-it` | 262K | | | | | | $0.08 | $0.30 |
145
144
  | `kilo/google/lyria-3-clip-preview` | 1.0M | | | | | | — | — |
146
145
  | `kilo/google/lyria-3-pro-preview` | 1.0M | | | | | | — | — |
147
146
  | `kilo/gryphe/mythomax-l2-13b` | 4K | | | | | | $0.08 | $0.11 |
@@ -4,7 +4,7 @@
4
4
 
5
5
  # ![NanoGPT logo](https://models.dev/logos/nano-gpt.svg)NanoGPT
6
6
 
7
- Access 595 NanoGPT models through Mastra's model router. Authentication is handled automatically using the `NANO_GPT_API_KEY` environment variable.
7
+ Access 597 NanoGPT models through Mastra's model router. Authentication is handled automatically using the `NANO_GPT_API_KEY` environment variable.
8
8
 
9
9
  Learn more in the [NanoGPT documentation](https://docs.nano-gpt.com).
10
10
 
@@ -595,6 +595,7 @@ for await (const chunk of stream) {
595
595
  | `nano-gpt/xiaomi/mimo-v2.5:thinking` | 1.0M | | | | | | $0.14 | $0.28 |
596
596
  | `nano-gpt/xiaomi/mimo-v2.6-flash` | 1.0M | | | | | | $0.14 | $0.28 |
597
597
  | `nano-gpt/xiaomi/mimo-v2.6-flash-uncensored` | 1.0M | | | | | | $0.50 | $2 |
598
+ | `nano-gpt/xiaomi/mimo-v2.6-flash-uncensored:thinking` | 1.0M | | | | | | $0.50 | $2 |
598
599
  | `nano-gpt/xiaomi/mimo-v2.6-pro` | 1.0M | | | | | | $0.43 | $0.87 |
599
600
  | `nano-gpt/xiaomi/mimo-v2.6-pro-ultraspeed` | 1.0M | | | | | | $4 | $9 |
600
601
  | `nano-gpt/z-ai/glm-4.5` | 128K | | | | | | $0.30 | $1 |
@@ -629,6 +630,7 @@ for await (const chunk of stream) {
629
630
  | `nano-gpt/z-ai/glm-5.3-flash` | 1.0M | | | | | | $0.07 | $0.25 |
630
631
  | `nano-gpt/z-ai/glm-5.3-flash-cybersecurity` | 1.0M | | | | | | $0.15 | $0.50 |
631
632
  | `nano-gpt/z-ai/glm-5.3-flash-uncensored` | 1.0M | | | | | | $0.20 | $0.80 |
633
+ | `nano-gpt/z-ai/glm-5.3-uncensored` | 1.0M | | | | | | $1 | $2 |
632
634
  | `nano-gpt/z-ai/glm-5.3:thinking` | 1.0M | | | | | | $1 | $3 |
633
635
  | `nano-gpt/z-ai/glm-5v-turbo` | 203K | | | | | | $1 | $4 |
634
636
  | `nano-gpt/z-ai/glm-5v-turbo:thinking` | 203K | | | | | | $1 | $4 |
@@ -4,7 +4,7 @@
4
4
 
5
5
  # ![OpenCode Go logo](https://models.dev/logos/opencode-go.svg)OpenCode Go
6
6
 
7
- Access 41 OpenCode Go models through Mastra's model router. Authentication is handled automatically using the `OPENCODE_API_KEY` environment variable.
7
+ Access 42 OpenCode Go models through Mastra's model router. Authentication is handled automatically using the `OPENCODE_API_KEY` environment variable.
8
8
 
9
9
  Learn more in the [OpenCode Go documentation](https://opencode.ai/docs/go).
10
10
 
@@ -56,6 +56,7 @@ for await (const chunk of stream) {
56
56
  | `opencode-go/kimi-k2.7-code` | 262K | | | | | | $0.95 | $4 |
57
57
  | `opencode-go/kimi-k3` | 1.0M | | | | | | $3 | $15 |
58
58
  | `opencode-go/longcat-2.0` | 1.0M | | | | | | $0.30 | $1 |
59
+ | `opencode-go/longcat-2.5-preview-free` | 1.0M | | | | | | — | — |
59
60
  | `opencode-go/mimo-v2.5` | 1.0M | | | | | | $0.14 | $0.28 |
60
61
  | `opencode-go/mimo-v2.5-pro` | 1.0M | | | | | | $0.43 | $0.87 |
61
62
  | `opencode-go/mimo-v2.6-flash` | 1.0M | | | | | | $0.14 | $0.28 |
@@ -4,7 +4,7 @@
4
4
 
5
5
  # ![OpenCode Zen logo](https://models.dev/logos/opencode.svg)OpenCode Zen
6
6
 
7
- Access 111 OpenCode Zen models through Mastra's model router. Authentication is handled automatically using the `OPENCODE_API_KEY` environment variable.
7
+ Access 112 OpenCode Zen models through Mastra's model router. Authentication is handled automatically using the `OPENCODE_API_KEY` environment variable.
8
8
 
9
9
  Learn more in the [OpenCode Zen documentation](https://opencode.ai/docs/zen).
10
10
 
@@ -100,6 +100,7 @@ for await (const chunk of stream) {
100
100
  | `opencode/kimi-k2.7-code` | 262K | | | | | | $0.95 | $4 |
101
101
  | `opencode/kimi-k3` | 1.0M | | | | | | $3 | $15 |
102
102
  | `opencode/ling-3.0-flash-fin-free` | 262K | | | | | | — | — |
103
+ | `opencode/longcat-2.5-preview-free` | 1.0M | | | | | | — | — |
103
104
  | `opencode/mimo-v2.6-flash-free` | 200K | | | | | | — | — |
104
105
  | `opencode/minimax-m2.5` | 205K | | | | | | $0.30 | $1 |
105
106
  | `opencode/minimax-m2.7` | 205K | | | | | | $0.30 | $1 |
@@ -54,8 +54,8 @@ for await (const chunk of stream) {
54
54
  | `requesty/claude-opus-4-8` | 1.0M | | | | | | $5 | $25 |
55
55
  | `requesty/claude-opus-4-8@eu` | 1.0M | | | | | | $6 | $28 |
56
56
  | `requesty/claude-opus-5` | 1.0M | | | | | | $5 | $25 |
57
- | `requesty/claude-opus-5-5` | 1.0M | | | | | | $5 | $25 |
58
- | `requesty/claude-opus-5-5@eu` | 1.0M | | | | | | $6 | $28 |
57
+ | `requesty/claude-opus-5-5` | 1.0M | | | | | | $4 | $20 |
58
+ | `requesty/claude-opus-5-5@eu` | 1.0M | | | | | | $4 | $22 |
59
59
  | `requesty/claude-opus-5@eu` | 1.0M | | | | | | $6 | $28 |
60
60
  | `requesty/claude-sonnet-4-5` | 1.0M | | | | | | $3 | $15 |
61
61
  | `requesty/claude-sonnet-4-5@eu` | 1.0M | | | | | | $3 | $17 |
@@ -268,6 +268,8 @@ Returns: `boolean`. `false` when this process has no active run recorded for the
268
268
 
269
269
  ### Recovery
270
270
 
271
+ > **Warning:** Recovery re-drives a run from its last persisted snapshot. Mastra treats a sub-agent delegation as one generated `agent-<name>` tool step and doesn't checkpoint its inner LLM or tool calls independently. If recovery replays that step, the entire sub-agent run can execute again, including already-completed tool side effects. Make tools used by sub-agents idempotent, or deduplicate their external side effects. See [crash recovery](https://mastra.ai/docs/harness/durable-agents) for operational guidance.
272
+
271
273
  #### `listActiveRuns(options?)`
272
274
 
273
275
  Lists this agent's runs whose persisted snapshot is in `running` status: runs whose agentic loop was mid-execution when the workflow engine last saved state. On a live process they transition to `suspended` or a terminal status. After a crash or restart they stay `running` with nothing driving them, which is what `recoverActiveRuns()` re-drives. Runs started by other durable agents on the same storage aren't included.
package/package.json CHANGED
@@ -1,6 +1,6 @@
1
1
  {
2
2
  "name": "@mastra/mcp-docs-server",
3
- "version": "1.3.2-alpha.6",
3
+ "version": "1.3.2-alpha.8",
4
4
  "description": "MCP server for accessing Mastra.ai documentation, changelogs, and news.",
5
5
  "type": "module",
6
6
  "main": "dist/index.js",
@@ -26,8 +26,8 @@
26
26
  "@mastra/mcp-legacy": "npm:@mastra/mcp@^1.18.0",
27
27
  "local-pkg": "^1.1.2",
28
28
  "zod": "^4.6.4",
29
- "@mastra/core": "1.72.0-alpha.3",
30
- "@mastra/mcp": "^2.1.0"
29
+ "@mastra/mcp": "^2.1.0",
30
+ "@mastra/core": "1.72.0-alpha.4"
31
31
  },
32
32
  "devDependencies": {
33
33
  "@hono/node-server": "^2.0.0",
@@ -44,8 +44,8 @@
44
44
  "typescript": "^7.0.2",
45
45
  "vitest": "4.1.11",
46
46
  "@internal/lint": "0.0.137",
47
- "@internal/types-builder": "0.0.112",
48
- "@mastra/core": "1.72.0-alpha.3"
47
+ "@mastra/core": "1.72.0-alpha.4",
48
+ "@internal/types-builder": "0.0.112"
49
49
  },
50
50
  "homepage": "https://mastra.ai",
51
51
  "repository": {