ur-agent 1.82.1 → 1.82.2

This diff represents the content of publicly available package versions that have been released to one of the supported registries. The information contained in this diff is provided for informational purposes only and reflects changes between package versions as they appear in their respective public registries.
package/CHANGELOG.md CHANGED
@@ -1,5 +1,22 @@
1
1
  # Changelog
2
2
 
3
+ ## 1.82.2
4
+
5
+ - Saved custom base URLs independently for every provider, including Ollama,
6
+ llama.cpp, vLLM, LM Studio, and Unsloth. Switching through `/provider`,
7
+ `/model`, or CLI configuration now restores each provider's last endpoint,
8
+ with automatic migration from the legacy single-address setting.
9
+ - Fixed Ollama graded reasoning discovery and transport. The model picker now
10
+ reads the focused model's `/api/show` capability before selection, Left/Right
11
+ cycles its real effort ladder, and requests send the chosen native `think`
12
+ value; Kimi K3 supports `low`, `high`, and `max`, while GPT-OSS safely caps at
13
+ `high`.
14
+ - Audited direct OpenAI, Anthropic, Gemini, OpenRouter, llama.cpp, and generic
15
+ OpenAI-compatible provider effort handling. Live model discovery retains
16
+ documented model-specific ladders, Gemini serializes `thinkingLevel`, and
17
+ providers without authoritative graded metadata no longer receive invented
18
+ levels.
19
+
3
20
  ## 1.82.1
4
21
 
5
22
  - Added a mandatory execution/trust kernel to every main, custom, and subagent
package/README.md CHANGED
@@ -295,6 +295,12 @@ ur config set provider.fallback ollama
295
295
  guidance. UR never switches providers automatically: inspect the failure and
296
296
  select the recovery provider explicitly with `ur config set provider <id>`.
297
297
 
298
+ `base_url` belongs to the provider that is active when it is set. UR remembers
299
+ each provider's address independently, so switching among Ollama, LM Studio,
300
+ llama.cpp, vLLM, Unsloth, or another compatible endpoint restores that
301
+ provider's last URL automatically. Existing single-URL settings are migrated
302
+ to the previously active provider on the first switch.
303
+
298
304
  OpenAI API traffic uses Chat Completions by default. The Responses transport is
299
305
  an explicit, provider-scoped opt-in with privacy-conscious defaults:
300
306
 
@@ -367,7 +373,9 @@ In the interactive app, `/model` is a two-step, provider-first picker:
367
373
  resolved visibly to that model's actual `max`, `xhigh`, or `high` ceiling,
368
374
  and the resolved value is the value sent to the provider. llama.cpp models
369
375
  are checked lazily through their model-scoped `/props` capability while the
370
- cursor moves. OpenRouter additionally
376
+ cursor moves. Ollama models are checked through `/api/show`; Kimi K3 exposes
377
+ `low`, `high`, and `max`, and the chosen level is sent through Ollama's
378
+ native `think` field. OpenRouter additionally
371
379
  shows compact model names, FREE/PAID tier, context size, tool/reasoning
372
380
  support, and the full untruncated ID immediately below the focused entry.
373
381
  Its catalog is fetched fresh every time it opens and never falls back to a