ur-agent 1.82.1 → 1.82.2
This diff represents the content of publicly available package versions that have been released to one of the supported registries. The information contained in this diff is provided for informational purposes only and reflects changes between package versions as they appear in their respective public registries.
- package/CHANGELOG.md +17 -0
- package/README.md +9 -1
- package/dist/cli.js +279 -164
- package/docs/CONFIGURATION.md +8 -0
- package/docs/TROUBLESHOOTING.md +4 -0
- package/docs/USAGE.md +5 -0
- package/docs/VALIDATION.md +1 -1
- package/docs/providers.md +16 -0
- package/documentation/index.html +1 -1
- package/extensions/jetbrains-ur/build.gradle.kts +1 -1
- package/extensions/vscode-ur-inline-diffs/package.json +1 -1
- package/package.json +1 -1
package/docs/CONFIGURATION.md
CHANGED
|
@@ -144,6 +144,11 @@ incompatible saved model instead of silently carrying it across providers. The
|
|
|
144
144
|
saved provider/model pair controls the runtime backend for the next agent
|
|
145
145
|
request; Ollama is only used when `ollama` is the selected provider.
|
|
146
146
|
|
|
147
|
+
The configured `base_url` is provider-scoped. Setting an address while vLLM is
|
|
148
|
+
active does not replace the saved Ollama, llama.cpp, or Unsloth address;
|
|
149
|
+
returning to any provider restores its own URL. Legacy `provider.baseUrl`
|
|
150
|
+
settings are migrated to the old active provider when the first switch occurs.
|
|
151
|
+
|
|
147
152
|
In the model step, Up/Down browses, Left/Right changes the focused model's
|
|
148
153
|
supported effort level, Enter confirms, Ctrl+R refreshes the catalog, and Esc
|
|
149
154
|
returns to providers. OpenRouter entries show pricing tier, context size,
|
|
@@ -151,6 +156,9 @@ tool/reasoning capability, compact names, and the exact ID for the focused
|
|
|
151
156
|
entry. Its catalog is fetched fresh whenever opened; a failed refresh never
|
|
152
157
|
silently displays cached entries. API-provider secret
|
|
153
158
|
entry stays on one masked row and stores the value through the keychain flow.
|
|
159
|
+
Ollama and llama.cpp capabilities are loaded lazily for the focused model from
|
|
160
|
+
`/api/show` and `/props`, respectively, so the arrow selector reflects the
|
|
161
|
+
actual model rather than a provider-wide guess.
|
|
154
162
|
|
|
155
163
|
The same provider-first picker is mandatory on the first interactive run in a
|
|
156
164
|
workspace with no model in `.ur/settings.json` or `.ur/settings.local.json`.
|
package/docs/TROUBLESHOOTING.md
CHANGED
|
@@ -231,6 +231,10 @@ ur config set base_url http://localhost:11434
|
|
|
231
231
|
ur provider doctor
|
|
232
232
|
```
|
|
233
233
|
|
|
234
|
+
Addresses are saved per provider. If the doctor probes an unexpected URL,
|
|
235
|
+
select that provider first and run `ur config get base_url`; changing vLLM's
|
|
236
|
+
address no longer overwrites Ollama, llama.cpp, or Unsloth.
|
|
237
|
+
|
|
234
238
|
## Sessions and workflows
|
|
235
239
|
|
|
236
240
|
### No visible progress in scripts
|
package/docs/USAGE.md
CHANGED
|
@@ -222,6 +222,11 @@ Unsloth defaults to `http://localhost:8888/v1`, requires its Studio API key,
|
|
|
222
222
|
and is inference-only: UR does not manage Unsloth and disables its server-side
|
|
223
223
|
tools while retaining standard function calls inside UR's guarded tool loop.
|
|
224
224
|
|
|
225
|
+
UR stores `base_url` per provider. You can set different addresses for
|
|
226
|
+
Ollama, llama.cpp, vLLM, and Unsloth once, then switch providers without
|
|
227
|
+
re-entering any of them. `ur config get base_url` always reports the address
|
|
228
|
+
for the currently active provider.
|
|
229
|
+
|
|
225
230
|
Use `/model` in an interactive session to select provider first and model
|
|
226
231
|
second. OpenAI API, Claude API, Gemini API, OpenRouter, Ollama, and
|
|
227
232
|
OpenAI-compatible endpoints stay separate; a subscription login does not grant
|
package/docs/VALIDATION.md
CHANGED
package/docs/providers.md
CHANGED
|
@@ -138,6 +138,11 @@ The fallback setting is a recovery hint for `ur provider doctor`; it does not
|
|
|
138
138
|
route a failed request to another provider. Review the failure and use
|
|
139
139
|
`ur config set provider <id>` to switch explicitly.
|
|
140
140
|
|
|
141
|
+
`base_url` is stored for the provider that is active when the command runs.
|
|
142
|
+
Each provider retains its own address across `/provider`, `/model`, and CLI
|
|
143
|
+
switches. The legacy single `provider.baseUrl` field remains readable and is
|
|
144
|
+
migrated to the previously active provider on the first switch.
|
|
145
|
+
|
|
141
146
|
OpenAI API uses Chat Completions by default. `openai_transport responses` is an
|
|
142
147
|
explicit opt-in to the native Responses adapter; it defaults to `store=false`
|
|
143
148
|
and supports semantic streaming, background polling/cancellation, WebSocket
|
|
@@ -167,6 +172,17 @@ indicator, active-work spinner, SDK settings response, and provider request all
|
|
|
167
172
|
use the same resolved value. If a provider advertises only boolean thinking,
|
|
168
173
|
UR does not invent a graded effort selector.
|
|
169
174
|
|
|
175
|
+
For Ollama, UR lazily reads the focused model's `/api/show` capabilities and
|
|
176
|
+
sends the selected level through native `think`. Kimi K3 uses
|
|
177
|
+
`low|high|max`; GPT-OSS uses `low|medium|high`; other models advertising
|
|
178
|
+
`thinking` use Ollama's current `low|medium|high|max` contract. Direct OpenAI,
|
|
179
|
+
Anthropic, and Gemini models use curated model-specific ladders from their
|
|
180
|
+
official documentation; live discovery rows are merged with those contracts.
|
|
181
|
+
See [Ollama thinking](https://docs.ollama.com/capabilities/thinking),
|
|
182
|
+
[OpenAI model guidance](https://developers.openai.com/api/docs/guides/latest-model),
|
|
183
|
+
[Claude effort](https://platform.claude.com/docs/en/build-with-claude/effort),
|
|
184
|
+
and [Gemini thinking](https://ai.google.dev/gemini-api/docs/thinking).
|
|
185
|
+
|
|
170
186
|
The provider-first `/model` picker supports the same control directly: use
|
|
171
187
|
Left/Right to move through the effort levels advertised by the focused model,
|
|
172
188
|
then Enter to apply the model and effort together. OpenRouter's live catalog
|
package/documentation/index.html
CHANGED
|
@@ -45,7 +45,7 @@
|
|
|
45
45
|
<main id="content" class="content">
|
|
46
46
|
<header class="topbar">
|
|
47
47
|
<div>
|
|
48
|
-
<p class="eyebrow">Version 1.82.
|
|
48
|
+
<p class="eyebrow">Version 1.82.2</p>
|
|
49
49
|
<h1>UR-Nexus Documentation</h1>
|
|
50
50
|
<p class="lead">A practical, tutorial-style reference for installing, configuring, automating, extending, and operating UR-Nexus.</p>
|
|
51
51
|
</div>
|
|
@@ -2,7 +2,7 @@
|
|
|
2
2
|
"name": "ur-inline-diffs",
|
|
3
3
|
"displayName": "UR Inline Diffs",
|
|
4
4
|
"description": "Review, apply, and reject UR inline diff bundles from .ur/ide/diffs inside VS Code.",
|
|
5
|
-
"version": "1.82.
|
|
5
|
+
"version": "1.82.2",
|
|
6
6
|
"publisher": "ur-nexus",
|
|
7
7
|
"engines": {
|
|
8
8
|
"vscode": "^1.92.0"
|
package/package.json
CHANGED