rigorrun 0.3.0 → 0.3.1

This diff represents the content of publicly available package versions that have been released to one of the supported registries. The information contained in this diff is provided for informational purposes only and reflects changes between package versions as they appear in their respective public registries.
package/README.md CHANGED
@@ -7,18 +7,22 @@
7
7
  Show RigorRun a job once. It turns that into a repeatable acceptance suite and decides whether your
8
8
  agent is safe to ship by reading the system it changed — never by trusting what it says about itself.
9
9
 
10
- **Early Access · v0.2** — parts of it are honestly unfinished, and they are listed rather than hidden.
10
+ **Early Access · v0.3** — parts of it are honestly unfinished, and they are listed rather than hidden.
11
11
 
12
12
  [rigorrun.xyz](https://rigorrun.xyz) · [Documentation](https://docs.rigorrun.xyz) · [Evidence](https://rigorrun.xyz/evidence) · [What is and is not built](https://rigorrun.xyz/what-is-built)
13
13
 
14
14
  </div>
15
15
 
16
16
  ```bash
17
- npx rigorrun
17
+ npx rigorrun demo # a real recorded run, replayed offline in a second
18
+ npx rigorrun # your own system and agent, in the local interface
18
19
  ```
19
20
 
20
- Open the URL it prints. Everything runs on your machine: there is no account, and no hosted
21
- component to send your systems to.
21
+ `demo` replays a real model working a bundled support desk: what it said beside what the system held
22
+ afterwards. The same run, case by case, is at [rigorrun.xyz/replay](https://rigorrun.xyz/replay).
23
+
24
+ `npx rigorrun` opens a local interface. Everything runs on your machine: there is no account, and no
25
+ hosted component to send your systems to.
22
26
 
23
27
  ---
24
28
 
@@ -45,7 +49,9 @@ most tools: you write the tests → the tool runs them
45
49
  be.
46
50
  3. **Rule on what it worked out.** It shows the evidence behind each proposed rule. A rule you
47
51
  reject cannot fail your agent.
48
- 4. **Connect your agent** — an HTTP endpoint, a local command, or your own loop pulling work.
52
+ 4. **Connect your agent** — unchanged, wherever it runs: RigorRun sends each case's work to a URL it
53
+ already serves and reads the result from your system (a black box). Or an HTTP endpoint, a local
54
+ command, or your own loop pulling work.
49
55
  5. **Run it.** A verdict, with how strongly each answer could be verified.
50
56
  6. **Gate the next change** in CI.
51
57