zuplo 7.8.14 → 7.8.15

This diff represents the content of publicly available package versions that have been released to one of the supported registries. The information contained in this diff is provided for informational purposes only and reflects changes between package versions as they appear in their respective public registries.
@@ -59,8 +59,8 @@ Mantle serves two API formats on the same host with the same API key:
59
59
  The gateway's model catalog records which API serves each model, and the gateway
60
60
  routes every request accordingly—your clients always call your app's URL and
61
61
  never see the Mantle endpoint. This also covers a Mantle quirk: AWS serves some
62
- models (the Gemma 4 and GPT-5.x families among them) on a different base path
63
- (`/openai/v1` instead of `/v1`), documented per
62
+ models (the Gemma 4, GPT-5.x, GPT-6, and Grok families among them) on a
63
+ different base path (`/openai/v1` instead of `/v1`), documented per
64
64
  [AWS model card](https://docs.aws.amazon.com/bedrock/latest/userguide/model-cards.html)—the
65
65
  per-model reference pages in the AWS Bedrock docs. The gateway sends each model
66
66
  to its documented path, so the model reference and your app's URL stay the same
@@ -227,12 +227,16 @@ Use the inference profile ID—usually the model ID with a regional prefix such
227
227
 
228
228
  The gateway prices Bedrock Runtime requests from the model catalog by exact
229
229
  model ID, the same as every other provider. Cross-region inference profiles are
230
- priced too: the `us.`, `eu.`, `apac.`, `au.`, and `global.` IDs are catalog rows
231
- in their own right. A model with no catalog row, such as a
232
- provisioned-throughput ARN or an inference profile AWS added after the catalog,
233
- is still served and still counts toward token and request
234
- [budgets](./usage-limits.mdx), but is priced at zero. It doesn't count toward
235
- spending budgets, and its response carries no `X-Cost-USD` header.
230
+ priced too: the `us.`, `eu.`, `apac.`, `au.`, `jp.`, `us-gov.`, and `global.`
231
+ IDs are catalog rows in their own right, each at its own rate. A
232
+ foundation-model or inference-profile ARN, such as
233
+ `arn:aws:bedrock:eu-central-1:123456789012:inference-profile/eu.anthropic.claude-opus-4-8`,
234
+ is priced as the ID it contains. A model with no catalog row, such as a
235
+ provisioned-throughput ARN, an application inference profile, or an inference
236
+ profile AWS added after the catalog, is still served and still counts toward
237
+ token and request [budgets](./usage-limits.mdx), but is priced at zero. It
238
+ doesn't count toward spending budgets, and its response carries no `X-Cost-USD`
239
+ header.
236
240
 
237
241
  When a budget is exhausted, these routes answer
238
242
  `400 ServiceQuotaExceededException`—not the `429` the Universal API returns. AWS
package/package.json CHANGED
@@ -1,6 +1,6 @@
1
1
  {
2
2
  "name": "zuplo",
3
- "version": "7.8.14",
3
+ "version": "7.8.15",
4
4
  "type": "module",
5
5
  "description": "The official Zuplo CLI for local development and platform management",
6
6
  "homepage": "https://zuplo.com/docs/cli/overview",
@@ -32,9 +32,9 @@
32
32
  "zuplo": "zuplo.js"
33
33
  },
34
34
  "dependencies": {
35
- "@zuplo/cli": "7.8.14",
36
- "@zuplo/core": "7.8.14",
37
- "@zuplo/runtime": "7.8.14",
38
- "@zuplo/test": "7.8.14"
35
+ "@zuplo/cli": "7.8.15",
36
+ "@zuplo/core": "7.8.15",
37
+ "@zuplo/runtime": "7.8.15",
38
+ "@zuplo/test": "7.8.15"
39
39
  }
40
40
  }