flex-race 0.1.0 → 0.1.1

This diff represents the content of publicly available package versions that have been released to one of the supported registries. The information contained in this diff is provided for informational purposes only and reflects changes between package versions as they appear in their respective public registries.
Files changed (2) hide show
  1. package/README.md +2 -0
  2. package/package.json +1 -1
package/README.md CHANGED
@@ -1,5 +1,7 @@
1
1
  # flex-race
2
2
 
3
+ For agents and coding tools, or to race Gemini and Claude models too, use the hosted version at [flexinference.com](https://flexinference.com). It speaks the OpenAI, Anthropic, and Gemini APIs, so an agent points its base URL at it with no code change. It races Gemini's flex tier and Anthropic's batch tier, and both tiers are priced at half the standard rate. Claude requests race only on FlexInference managed keys, without streaming, and only when the deadline allows three to ten minutes.
4
+
3
5
  flex-race lets your app try OpenAI's flex tier with a set wait for it to start. If flex takes too long or fails before it starts, the wrapper sends the request to the default tier. This gives your app a way to use flex while limiting the wait for flex to accept the work. The [wrapper](src/with-flex.ts) adds start_by to responses.create on an OpenAI client.
4
6
 
5
7
  ## Install
package/package.json CHANGED
@@ -1,6 +1,6 @@
1
1
  {
2
2
  "name": "flex-race",
3
- "version": "0.1.0",
3
+ "version": "0.1.1",
4
4
  "description": "Race OpenAI's flex tier against a deadline on the Responses API, falling back to the default tier.",
5
5
  "license": "MIT",
6
6
  "author": "Aditya Perswal",