orchestrator-workflow 0.40.0 → 0.40.1

This diff represents the content of publicly available package versions that have been released to one of the supported registries. The information contained in this diff is provided for informational purposes only and reflects changes between package versions as they appear in their respective public registries.
package/CHANGELOG.md CHANGED
@@ -7,6 +7,22 @@ and this project adheres to [Semantic Versioning](https://semver.org/spec/v2.0.0
7
7
 
8
8
  ## [Unreleased]
9
9
 
10
+ ## [0.40.1] - 2026-09-24
11
+
12
+ - Restored the `0.37.0` section byte for byte to the text released at tag
13
+ `orchestrator-workflow/v0.37.0`, after a later change had rewritten its
14
+ wording. The test assertion that read that section was removed; the rule
15
+ clauses stay pinned against the reference files and the docs/okf bundle
16
+ doc, since a released section stays as shipped.
17
+
18
+ - The shared verification-set comparison sentence in implementer.md,
19
+ reviewer.md, and contracts.md now compares effective config and scripts,
20
+ and preflight executable identity/definition, at the tree the
21
+ verification set executes in, and defers repository identity to the
22
+ path rule instead of restating it against the role's own checkout. The
23
+ two rules read as a single rule again instead of two that could be read
24
+ as disagreeing.
25
+
10
26
  ## [0.40.0] - 2026-09-23
11
27
 
12
28
  - A verification set can no longer look green while checking the wrong
@@ -316,27 +332,27 @@ and this project adheres to [Semantic Versioning](https://semver.org/spec/v2.0.0
316
332
  the check itself unchanged is a `03-decisions.md` entry, not a revision";
317
333
  the orchestrator records that entry, states in it why no evidence is
318
334
  invalidated, and communicates the corrected wording in the next
319
- delegation. Step 7 now says: "For a review round whose entire delta contains only explanatory documentation, comments, or citations and contains no source- or test-file edits and no semantic change to executable commands, configuration, policy, instructions, or behavior, default to the `-medium` reviewer tier with `review_method: normal` where tier variants are installed". That refines the general tier default for this one class only. "A review round that touches an instruction, policy, template or prompt file (for example a SKILL.md instruction) keeps the general default, whatever the file type, and the minimums named above are unaffected";
335
+ delegation. Step 7 now says: "For a review round whose entire delta is a
336
+ docs-only delta in the sense of step 8's docs-only closure, default to the
337
+ `-medium` reviewer tier with `review_method: normal` where tier variants
338
+ are installed". That refines the general tier default for this one class
339
+ only, a round that touches an instruction, policy, template or prompt file
340
+ keeps the general default, and the minimum review methods are untouched;
320
341
  the AGENTS.md section does not yet point to this refinement.
321
342
  `references/review-and-recovery.md` gains a "Pinned-prose changes" section
322
343
  for a change whose acceptance rests on tests that pin documentation
323
344
  wording: "A prose mutant survives exactly when its bytes sit in no
324
345
  assertion", so review rounds that hunt for the next unpinned sentence do
325
- not converge. The section asks for one normative site per rule, a claim list in the acceptance criterion as the pin obligation. "Every normative sentence the change adds or alters at that site is pinned; one left unpinned is named in the criterion with the reason it is not load-bearing." A reviewer briefing
326
- bounds the prose mutant space to that list, and "When the briefing
327
- bounds the
328
- prose mutant space to a claim list, respect that bound and put scope notes in
329
- `residual_risks`, unless an unlisted sentence is shown to be load-bearing."
330
- Copies are bound to the normative site by one shared test constant, and: "Cap
331
- test-adequacy review rounds on the change at two." "A test-adequacy review
332
- round is one whose returned findings are all `tests` findings of severity
333
- `low` or `medium` about pin gaps on the pinned prose; a round returning any
334
- other finding is an ordinary round outside the cap." "The cap changes neither
335
- the Round-2 halt rule, the Review-round escalation budget nor the
336
- Fix-regression decision point: a test-adequacy review round still counts as a
337
- negative round where it is one." It exempts semantic findings and leaves the
338
- review gate as it is. Step 7 points to the section without
339
- restating it. Evidence (issue #300 and the change that added the Fix-regression
346
+ not converge. The section asks for one normative site per rule, a claim
347
+ list in the acceptance criterion as the pin obligation (every normative
348
+ sentence the change adds or alters at that site is a claim, an omission is
349
+ named with its reason), a reviewer briefing that bounds the prose mutant
350
+ space to that list, copies bound to the normative site by one shared test
351
+ constant, and: "Cap test-adequacy review rounds on the change at two." It
352
+ defines the capped round, exempts semantic findings, and changes neither
353
+ the Round-2 halt rule, the escalation budget, the Fix-regression decision
354
+ point nor the review gate. Step 7 points to the section without restating
355
+ it. Evidence (issue #300 and the change that added the Fix-regression
340
356
  decision point; one repository each, not a benchmark): the issue reports
341
357
  baseline revisions r1 to r3 for two wording precisions of a verification
342
358
  method, and a run in which the top reviewer tier was about half the day's
@@ -44,11 +44,11 @@ Rules:
44
44
  orchestrator's explicit approval of the resolved repository configuration
45
45
  and every script/argument, since a repository set is not authority to
46
46
  execute repository data on its own. Before acquisition or execution,
47
- compare the frozen snapshot's effective config and scripts, preflight
48
- executable identity/definition, and repository identity with the tree the
49
- role runs in; any mismatch withdraws the approval like a digest mismatch
50
- and is reported as a misfire, and a change the task's own diff makes to
51
- one of those components is outside the approval. The compared values are
47
+ compare the frozen snapshot's effective config and scripts and preflight executable
48
+ identity/definition at the tree the set executes in; repository identity follows the
49
+ path rule, not the role's checkout; any mismatch withdraws the approval like a digest
50
+ mismatch and is reported as a misfire, and a change the task's own diff makes to one
51
+ of those components is outside the approval. The compared values are
52
52
  the ones recorded in the frozen snapshot at the run-local path
53
53
  `verification_set.snapshot` names; evidence-and-probes.md's Verification
54
54
  sets section defines what counts as a script for that comparison. Use the
@@ -61,10 +61,10 @@ Check, at minimum:
61
61
  outside it still requires confirming the orchestrator approved the resolved
62
62
  effective configuration and scripts, since a repository set is not
63
63
  authority to execute repository data on its own. Before acquisition or
64
- execution, compare the frozen snapshot's effective config and scripts,
65
- preflight executable identity/definition, and repository identity with the
66
- tree the role runs in; any mismatch withdraws the approval like a digest
67
- mismatch and is reported as a misfire, and a change the task's own diff
64
+ execution, compare the frozen snapshot's effective config and scripts and preflight
65
+ executable identity/definition at the tree the set executes in; repository identity
66
+ follows the path rule, not the role's checkout; any mismatch withdraws the approval
67
+ like a digest mismatch and is reported as a misfire, and a change the task's own diff
68
68
  makes to one of those components is outside the approval. The compared
69
69
  values are the ones recorded in the frozen snapshot at the run-local path
70
70
  `verification_set.snapshot` names; evidence-and-probes.md's Verification
@@ -86,11 +86,11 @@ approval reaches only the frozen
86
86
  snapshot: an unfrozen set, a changed script, or anything the snapshot does not
87
87
  capture still needs the orchestrator's explicit approval before acquisition or
88
88
  execution, since a repository set is not authority to execute repository data
89
- on its own. Before acquisition or execution, compare the frozen snapshot's
90
- effective config and scripts, preflight executable identity/definition, and
91
- repository identity with the tree the role runs in; any mismatch withdraws
92
- the approval like a digest mismatch and is reported as a misfire, and a
93
- change the task's own diff makes to one of those components is outside the
89
+ on its own. Before acquisition or execution, compare the frozen snapshot's effective
90
+ config and scripts and preflight executable identity/definition at the tree the set
91
+ executes in; repository identity follows the path rule, not the role's checkout; any
92
+ mismatch withdraws the approval like a digest mismatch and is reported as a misfire,
93
+ and a change the task's own diff makes to one of those components is outside the
94
94
  approval. The compared values are the ones recorded in the frozen snapshot at
95
95
  the run-local path `verification_set.snapshot` names; evidence-and-probes.md's
96
96
  Verification sets section defines what counts as a script for that
package/package.json CHANGED
@@ -1,6 +1,6 @@
1
1
  {
2
2
  "name": "orchestrator-workflow",
3
- "version": "0.40.0",
3
+ "version": "0.40.1",
4
4
  "description": "Installer for an orchestrator-led agent workflow: .ai/ run state, an AGENTS.md policy section, and per-harness subagent definitions for Claude Code, OpenAI Codex, and opencode",
5
5
  "main": "dist/index.js",
6
6
  "type": "module",