@codyswann/lisa 2.263.0 → 2.264.0

This diff represents the content of publicly available package versions that have been released to one of the supported registries. The information contained in this diff is provided for informational purposes only and reflects changes between package versions as they appear in their respective public registries.
Files changed (90) hide show
  1. package/dist/core/upstream-evidence-manifest.d.ts.map +1 -1
  2. package/dist/core/upstream-evidence-manifest.js +9 -7
  3. package/dist/core/upstream-evidence-manifest.js.map +1 -1
  4. package/package.json +1 -1
  5. package/plugins/lisa/.claude-plugin/plugin.json +1 -1
  6. package/plugins/lisa/.codex-plugin/plugin.json +1 -1
  7. package/plugins/lisa/.codex-plugin/skills/lisa-github-evidence/SKILL.md +10 -0
  8. package/plugins/lisa/.codex-plugin/skills/lisa-jira-evidence/SKILL.md +10 -0
  9. package/plugins/lisa/.codex-plugin/skills/lisa-linear-evidence/SKILL.md +10 -0
  10. package/plugins/lisa/.codex-plugin/skills/lisa-tracker-evidence/SKILL.md +10 -3
  11. package/plugins/lisa/rules/eager/claim-evidence-mapping.md +23 -3
  12. package/plugins/lisa/rules/reference/claim-evidence-mapping.md +134 -8
  13. package/plugins/lisa/rules/reference/verification.md +26 -0
  14. package/plugins/lisa/skills/lisa-github-evidence/SKILL.md +10 -0
  15. package/plugins/lisa/skills/lisa-jira-evidence/SKILL.md +10 -0
  16. package/plugins/lisa/skills/lisa-linear-evidence/SKILL.md +10 -0
  17. package/plugins/lisa/skills/lisa-tracker-evidence/SKILL.md +10 -3
  18. package/plugins/lisa-agy/plugin.json +1 -1
  19. package/plugins/lisa-agy/skills/lisa-github-evidence/SKILL.md +10 -0
  20. package/plugins/lisa-agy/skills/lisa-jira-evidence/SKILL.md +10 -0
  21. package/plugins/lisa-agy/skills/lisa-linear-evidence/SKILL.md +10 -0
  22. package/plugins/lisa-agy/skills/lisa-tracker-evidence/SKILL.md +10 -3
  23. package/plugins/lisa-cdk/.claude-plugin/plugin.json +1 -1
  24. package/plugins/lisa-cdk/.codex-plugin/plugin.json +1 -1
  25. package/plugins/lisa-cdk-agy/plugin.json +1 -1
  26. package/plugins/lisa-cdk-copilot/.claude-plugin/plugin.json +1 -1
  27. package/plugins/lisa-cdk-cursor/.claude-plugin/plugin.json +1 -1
  28. package/plugins/lisa-copilot/.claude-plugin/plugin.json +1 -1
  29. package/plugins/lisa-copilot/rules/eager/claim-evidence-mapping.md +23 -3
  30. package/plugins/lisa-copilot/rules/reference/claim-evidence-mapping.md +134 -8
  31. package/plugins/lisa-copilot/rules/reference/verification.md +26 -0
  32. package/plugins/lisa-copilot/skills/lisa-github-evidence/SKILL.md +10 -0
  33. package/plugins/lisa-copilot/skills/lisa-jira-evidence/SKILL.md +10 -0
  34. package/plugins/lisa-copilot/skills/lisa-linear-evidence/SKILL.md +10 -0
  35. package/plugins/lisa-copilot/skills/lisa-tracker-evidence/SKILL.md +10 -3
  36. package/plugins/lisa-cursor/.claude-plugin/plugin.json +1 -1
  37. package/plugins/lisa-cursor/rules/claim-evidence-mapping-reference.mdc +134 -8
  38. package/plugins/lisa-cursor/rules/claim-evidence-mapping.mdc +23 -3
  39. package/plugins/lisa-cursor/rules/verification-reference.mdc +26 -0
  40. package/plugins/lisa-cursor/skills/lisa-github-evidence/SKILL.md +10 -0
  41. package/plugins/lisa-cursor/skills/lisa-jira-evidence/SKILL.md +10 -0
  42. package/plugins/lisa-cursor/skills/lisa-linear-evidence/SKILL.md +10 -0
  43. package/plugins/lisa-cursor/skills/lisa-tracker-evidence/SKILL.md +10 -3
  44. package/plugins/lisa-expo/.claude-plugin/plugin.json +1 -1
  45. package/plugins/lisa-expo/.codex-plugin/plugin.json +1 -1
  46. package/plugins/lisa-expo-agy/plugin.json +1 -1
  47. package/plugins/lisa-expo-copilot/.claude-plugin/plugin.json +1 -1
  48. package/plugins/lisa-expo-cursor/.claude-plugin/plugin.json +1 -1
  49. package/plugins/lisa-harper-fabric/.claude-plugin/plugin.json +1 -1
  50. package/plugins/lisa-harper-fabric/.codex-plugin/plugin.json +1 -1
  51. package/plugins/lisa-harper-fabric-agy/plugin.json +1 -1
  52. package/plugins/lisa-harper-fabric-copilot/.claude-plugin/plugin.json +1 -1
  53. package/plugins/lisa-harper-fabric-cursor/.claude-plugin/plugin.json +1 -1
  54. package/plugins/lisa-nestjs/.claude-plugin/plugin.json +1 -1
  55. package/plugins/lisa-nestjs/.codex-plugin/plugin.json +1 -1
  56. package/plugins/lisa-nestjs-agy/plugin.json +1 -1
  57. package/plugins/lisa-nestjs-copilot/.claude-plugin/plugin.json +1 -1
  58. package/plugins/lisa-nestjs-cursor/.claude-plugin/plugin.json +1 -1
  59. package/plugins/lisa-openclaw/.claude-plugin/plugin.json +1 -1
  60. package/plugins/lisa-openclaw/.codex-plugin/plugin.json +1 -1
  61. package/plugins/lisa-openclaw-agy/plugin.json +1 -1
  62. package/plugins/lisa-openclaw-copilot/.claude-plugin/plugin.json +1 -1
  63. package/plugins/lisa-openclaw-cursor/.claude-plugin/plugin.json +1 -1
  64. package/plugins/lisa-phaser/.claude-plugin/plugin.json +1 -1
  65. package/plugins/lisa-phaser/.codex-plugin/plugin.json +1 -1
  66. package/plugins/lisa-phaser-agy/plugin.json +1 -1
  67. package/plugins/lisa-phaser-copilot/.claude-plugin/plugin.json +1 -1
  68. package/plugins/lisa-phaser-cursor/.claude-plugin/plugin.json +1 -1
  69. package/plugins/lisa-rails/.claude-plugin/plugin.json +1 -1
  70. package/plugins/lisa-rails/.codex-plugin/plugin.json +1 -1
  71. package/plugins/lisa-rails-agy/plugin.json +1 -1
  72. package/plugins/lisa-rails-copilot/.claude-plugin/plugin.json +1 -1
  73. package/plugins/lisa-rails-cursor/.claude-plugin/plugin.json +1 -1
  74. package/plugins/lisa-typescript/.claude-plugin/plugin.json +1 -1
  75. package/plugins/lisa-typescript/.codex-plugin/plugin.json +1 -1
  76. package/plugins/lisa-typescript-agy/plugin.json +1 -1
  77. package/plugins/lisa-typescript-copilot/.claude-plugin/plugin.json +1 -1
  78. package/plugins/lisa-typescript-cursor/.claude-plugin/plugin.json +1 -1
  79. package/plugins/lisa-wiki/.claude-plugin/plugin.json +1 -1
  80. package/plugins/lisa-wiki/.codex-plugin/plugin.json +1 -1
  81. package/plugins/lisa-wiki-agy/plugin.json +1 -1
  82. package/plugins/lisa-wiki-copilot/.claude-plugin/plugin.json +1 -1
  83. package/plugins/lisa-wiki-cursor/.claude-plugin/plugin.json +1 -1
  84. package/plugins/src/base/rules/eager/claim-evidence-mapping.md +23 -3
  85. package/plugins/src/base/rules/reference/claim-evidence-mapping.md +134 -8
  86. package/plugins/src/base/rules/reference/verification.md +26 -0
  87. package/plugins/src/base/skills/lisa-github-evidence/SKILL.md +10 -0
  88. package/plugins/src/base/skills/lisa-jira-evidence/SKILL.md +10 -0
  89. package/plugins/src/base/skills/lisa-linear-evidence/SKILL.md +10 -0
  90. package/plugins/src/base/skills/lisa-tracker-evidence/SKILL.md +10 -3
@@ -80,12 +80,12 @@ defines the names; it stores nothing:
80
80
  | `required_evidence_kinds` | the evidence kind(s) that reach that boundary, from the `verification` artifact-type set |
81
81
 
82
82
  A claim whose `required_evidence_kinds` has no captured, reaching artifact is **Not established** —
83
- the concept is named here and defined fully, with its evidence templates, in **BCE-3 (#1837)**; do
84
- not assume that section is present in this branch. Artifact identity what makes two captured
85
- artifacts the same or different is pinned in **BCE-4 (#1838)**, and the conservative default
86
- bucket for a security-sensitive claim is set in **BCE-5 (#1839)**. Each is named here as the field
87
- BCE-2's schema will carry; none is defined by this contract. Each ships with that ticket do not
88
- assume its section is present in this branch.
83
+ defined in full, with its evidence templates, in the section below. Artifact identity — what makes
84
+ two captured artifacts the same or different — is defined in the *Artifact identity* section below
85
+ (shipped by BCE-4, #1838). The conservative default bucket for a security-sensitive claim is set in
86
+ **BCE-5 (#1839)** named here as a field BCE-2's schema carries, not defined by this contract, and
87
+ it ships with that ticket. Where such a sibling surface is not installed in a given branch, name what
88
+ you can and continue.
89
89
 
90
90
  ### Worked example
91
91
 
@@ -114,13 +114,139 @@ Claim: "The service is deployed and healthy."
114
114
  response from the target environment.
115
115
  ```
116
116
 
117
+ ## Artifact identity — what the evidence was collected against
118
+
119
+ A claim reaches only as far as its evidence's *kind*; it applies only to the *artifact* that evidence
120
+ was collected against. **Artifact identity is what makes two captured artifacts the same or
121
+ different**: one repository, at one commit, in one environment, at one moment. Evidence that does not
122
+ say which artifact it observed is not evidence about anything in particular — it silently transfers
123
+ to whatever ships next, which is exactly how an auto-merge race ships code no verification ever
124
+ touched.
125
+
126
+ Identity is carried in two places, and they must agree:
127
+
128
+ | Where | Field | What it pins |
129
+ |---|---|---|
130
+ | `artifact` (once per verdict) | `repository` | the `owner/repo` the run observed |
131
+ | | `base_sha` | the base the change was measured against |
132
+ | | `head_sha` | **the commit the verification actually observed** — required for a v2 pass |
133
+ | | `build_id` | the build/run the evidence came from, where one exists |
134
+ | | `environment` | where it ran (local, preview, staging, production) |
135
+ | | `observed_at` | when the run observed it (ISO-8601 UTC) |
136
+ | `evidence[]` (per artifact) | `artifact_head_sha` | the `head_sha` in force **when that artifact was captured** |
137
+ | | `sha256` | content digest of the committed evidence file |
138
+ | | `captured_at` | when that artifact was captured (ISO-8601 UTC) |
139
+
140
+ `head_sha` pins the *build*; `sha256` pins the *bytes*. Together they answer both identity questions:
141
+ "which artifact was this collected against" and "is this still the artifact that was collected". The
142
+ `sha256` + commit-ref discipline is the same one Lisa's upstream-evidence manifest already uses — it
143
+ is cited as prior art here, not reinvented.
144
+
145
+ ### Two identity failures, both loud
146
+
147
+ - **`artifact_mismatch`** — an `evidence[]` entry whose `artifact_head_sha` differs from the verdict's
148
+ `artifact.head_sha`. The evidence describes a different build than the one the verdict claims. The
149
+ identity check fails **loudly, naming both SHAs** (the evidence's and the verdict's) so an operator
150
+ can see which build each half is talking about.
151
+ - **`evidence_digest_mismatch`** — on read, recompute the `sha256` of each committed evidence file. If
152
+ the bytes no longer match the recorded digest, the artifact has changed since it was recorded: the
153
+ check fails, it names the evidence id, and blocks completion. A digest that cannot be recomputed
154
+ (file absent) is the same failure.
155
+
156
+ Neither failure is ever summarized as "verification failed". Name the field, the ids, and both SHAs —
157
+ a person who does not code reads this at the gate (`factory-model` rule 5). Both checks are
158
+ **advisory-first**: reported to the operator but non-blocking until
159
+ `verification.gate.enforceBoundaries` is `true` in `.lisa.config.json` — the same ratchet flag the
160
+ boundary checks and the Stop-hook gate ride.
161
+
162
+ ### The merge race — one definition, two guards
163
+
164
+ "The artifact that shipped" is already defined once, in `lisa-drive-pr-to-merge`: the **ancestry**
165
+ check (is the verified commit an ancestor of the merged base branch, and is the merge commit's parent
166
+ that verified head rather than a stale one) **plus** the **deploy-run** check (a deploy/release run
167
+ actually fired for the merge SHA or an including descendant). Cite that definition; never write a
168
+ second one. Identity reconciliation is the same definition read from the evidence side:
169
+
170
+ - Evidence collected on a **pre-merge head** is valid for the merge commit **only when that head is a
171
+ parent of the merge** — i.e. it satisfies that skill's ancestry check against the merged head. If a
172
+ late commit raced past the merge, the merged head is not the verified one and the evidence does not
173
+ transfer.
174
+ - Ancestry alone is never enough. That skill's rule stands unchanged here: **never report shipped on
175
+ ancestry alone** — the deploy-run check must also pass before completion is declared.
176
+ - On mismatch (`artifact.head_sha` ≠ the reconciled merge SHA), completion is not declared. Flag the
177
+ mismatch naming both SHAs and **re-run verification against the merged head**; the re-run's verdict
178
+ is what may declare completion.
179
+
180
+ Two guards, one definition — so they cannot disagree about what shipped.
181
+
182
+ ## The "Not established" section — required, never omitted
183
+
184
+ Every report that asserts something was verified also states, in the same breath, **what it did not
185
+ establish**. This is one section, it is **required**, and it is **never omitted and never blank** —
186
+ on the evidence comment, in the committed `evidence/<ticket>/verdict.json`, and in the
187
+ `verification-status.json` verdict BCE-2's gate reads.
188
+
189
+ **Where it appears and what it says**
190
+
191
+ | Surface | Shape |
192
+ |---|---|
193
+ | Evidence comment (tracker + PR `## Evidence` section) | a `## Not established` heading, one plain-language bullet per item |
194
+ | `evidence/<ticket>/verdict.json` | `not_established: []` plus `not_established_reviewed: true` |
195
+ | `.lisa/verification-status.json` (schema v2) | per-claim `not_established[]` plus the top-level `not_established_reviewed` flag |
196
+
197
+ **The empty case is not the omitted case.** A verification that genuinely left nothing unproved still
198
+ renders the heading, with the single line:
199
+
200
+ ```text
201
+ ## Not established
202
+
203
+ None outstanding — reviewed
204
+ ```
205
+
206
+ An absent heading, or a heading with nothing under it, is a defect — it is indistinguishable from
207
+ never having asked the question. The list may be empty; the section may not be blank.
208
+
209
+ **Machine-readable semantics (as shipped by BCE-2, #1836).** `not_established` is the list — it may
210
+ be empty. `not_established_reviewed` is the boolean attestation that the list was actually
211
+ reviewed — **the flag may never be omitted**. That asymmetry is the whole mechanism: an empty list
212
+ plus a present flag means "we looked and found nothing outstanding"; an absent flag means nobody
213
+ looked. The Stop-hook gate treats an absent flag as a v2 contract violation, reported to stderr and
214
+ **advisory** until `verification.gate.enforceBoundaries` is `true` in `.lisa.config.json` (the same
215
+ ratchet flag the boundary checks ride); evidence surfaces refuse the post on the same terms.
216
+
217
+ **What belongs under the heading** — written in operator voice (`factory-model` rule 5: a person who
218
+ does not code reads this at the gate), not in engineering shorthand:
219
+
220
+ - **Boundaries not exercised** — a claim's boundary that no captured artifact reached. *"The checkout
221
+ button was proved in the browser; the order's persisted row was never queried, so the `data`
222
+ boundary is not established."*
223
+ - **Environments not tested** — where it was and was not run. *"Checked on production Chrome at
224
+ 1440×900 only. Not checked on mobile Safari, and not checked against the staging database."*
225
+ - **Claims consciously out of scope** — deliberately excluded behavior, named so nobody infers it.
226
+ *"Refunds were not touched or tested; this change covers new orders only."*
227
+ - **Anything a green quality check might be mistaken for proving.** *"Unit tests pass for the submit
228
+ handler. That establishes the code-unit boundary only — it is not evidence the button works."*
229
+
230
+ Each item names the thing, not a category: "not tested on mobile Safari" is usable at a gate; "some
231
+ environments untested" is not.
232
+
233
+ ### Philosophical precedent for this section
234
+
235
+ This generalizes `lisa-improve-harness`'s **`Known limits`** field — a required, never-empty line on
236
+ every result record. That skill says it plainly:
237
+ *"A record with nothing in it is invalid on its face"* — because a single-trajectory loop always has
238
+ limits. The same is true of any verification:
239
+ it ran somewhere, on something, once. `Known limits` (one record) and `not_established` (every claim
240
+ in the factory) are the same discipline with the same never-empty rule; read either and you should
241
+ recognize the other.
242
+
117
243
  ## Philosophical precedent
118
244
 
119
245
  This generalizes the **bounded-claim discipline** of `lisa-improve-harness`: one trajectory supports
120
246
  one trajectory's claim, and a result record may claim only what its cited evidence reaches. Here the
121
247
  same discipline is applied to every claim in the factory — a claim reaches exactly as far as the
122
- *kind* of evidence behind it, and no further. BCE-3 generalizes the *Not established* half of that
123
- discipline into a first-class report state.
248
+ *kind* of evidence behind it, and no further. The *Not established* half of that discipline is a
249
+ first-class report state — see the required, never-omitted section above.
124
250
 
125
251
  ## No behavior change; degrade, never block
126
252
 
@@ -151,6 +151,32 @@ Do not invent types inline; if none fits, propose extending this table. The lega
151
151
 
152
152
  The manifest is the single source of truth for "what evidence is required": authored once in the Validation Journey, enforced at write time, replayed during `tracker-journey` (which captures each artifact **in its declared type**), and checked again before the ticket closes. There is no second list to keep in sync.
153
153
 
154
+ ### Every evidence surface names what it did NOT establish
155
+
156
+ An evidence comment that lists only what passed is unreadable at a gate: a journey that skipped an edge state looks exactly like one that covered it. So every evidence comment — and the committed `evidence/<ticket>/verdict.json` — carries two extra sections, defined in full by the `claim-evidence-mapping` rule:
157
+
158
+ - **Artifact identity** — what the evidence was collected against, as values rather than placeholders: the `repository`, the `head_sha` the run observed, the `environment`, and per artifact its `sha256` digest and `captured_at`. Defined in full by the `claim-evidence-mapping` rule.
159
+ - **Not established** — a **required, never-omitted** heading listing what the verification did *not* prove: boundaries not exercised, environments not tested, behavior consciously out of scope. When nothing is outstanding it still renders, reading `None outstanding — reviewed`. It is never blank.
160
+
161
+ The committed verdict carries the machine-readable half:
162
+
163
+ ```
164
+ evidence/<ticket>/verdict.json
165
+ not_established: [] # what was NOT proved; may be empty
166
+ not_established_reviewed: true # attests the list was reviewed; may NEVER be omitted
167
+ artifact: { repository, base_sha, head_sha, build_id, environment, observed_at }
168
+ evidence: [ { evidence_id, kind, locator,
169
+ artifact_head_sha, # the head_sha in force when THIS artifact was captured
170
+ sha256, # content digest of the committed evidence file
171
+ captured_at } ]
172
+ ```
173
+
174
+ `artifact.head_sha` pins the build the verification observed; each entry's `sha256` pins the bytes. An entry whose `artifact_head_sha` differs from `artifact.head_sha` is an `artifact_mismatch` and a recomputed digest that disagrees is an `evidence_digest_mismatch` — each fails loudly, naming both SHAs or the evidence id, and blocks completion. At completion the pinned `head_sha` is reconciled against **the merged head** using the ancestry + deploy-run definition of "what shipped" that `lisa-drive-pr-to-merge` already owns (cite it; there is no second definition): pre-merge evidence counts only when its head is a parent of the merge, and on a merge-race mismatch verification re-runs against the merged head before completion is declared.
175
+
176
+ The list may be empty; the flag may not be missing. An absent `not_established_reviewed` is indistinguishable from nobody having asked the question, so the evidence-posting gate in `tracker-evidence` refuses the post, and the Stop-hook gate reports it as a v2 contract violation (advisory until `verification.gate.enforceBoundaries` is ratcheted on). This generalizes the required, never-empty `Known limits` field of `lisa-improve-harness` to every evidence surface.
177
+
178
+ The boundary each artifact type reaches — and therefore which claim a captured artifact can discharge — is the `claim-evidence-mapping` rule's taxonomy; the type table above is its evidence-kind source.
179
+
154
180
  ### Cross-work-item evidence references are non-claiming
155
181
 
156
182
  When prose needs to point at evidence declared by another work item, use the dedicated reference form:
@@ -24,6 +24,16 @@ Upload captured evidence and generated templates to the GitHub PR description an
24
24
  - `comment.md` — GitHub markdown body for both the issue comment and the PR description's `## Evidence` section.
25
25
  - (Optional) `comment.txt` — kept for parity with the JIRA path; not used here.
26
26
 
27
+ ## Comment-body preflight (required)
28
+
29
+ Before posting or updating anything, check the evidence body (`comment.md`, and `comment.txt` where this skill uses it):
30
+
31
+ - It contains a `## Not established` heading. That heading is **never omitted and never blank** — when nothing is outstanding it still renders `None outstanding — reviewed`; otherwise it names, in plain operator language, what the verification did not prove.
32
+ - The accompanying verdict carries `not_established_reviewed: true` (the list may be empty; the flag may never be omitted).
33
+ - It contains a `## Artifact identity` heading carrying **values, not placeholders** — the repository, the `head_sha` the verification observed, the `environment`, and per artifact its `sha256` digest and `captured_at`. **Refuse to post** a body whose identity heading is absent or unpopulated, or whose recorded `artifact_head_sha` disagrees with the verdict's `artifact.head_sha` — report the evidence id and **both SHAs**. Definition: the `claim-evidence-mapping` rule.
34
+
35
+ If either is missing, **refuse to post**: stop and report the missing Not-established review to the caller instead of publishing. Composing the body is `lisa-tracker-evidence`'s job (see its UI Evidence Checklist); this skill only refuses to publish one that omits the section. The section is defined by the `claim-evidence-mapping` rule and generalizes `lisa-improve-harness`'s required, never-empty `Known limits` field.
36
+
27
37
  ## Workflow
28
38
 
29
39
  1. **Resolve refs**
@@ -42,6 +42,16 @@ Upload captured evidence and generated templates to GitHub PR description and JI
42
42
  - `comment.txt` — JIRA wiki markup (generated by `generate-templates.py`)
43
43
  - `comment.md` — GitHub markdown (generated by `generate-templates.py`)
44
44
 
45
+ ## Comment-body preflight (required)
46
+
47
+ Before posting or updating anything, check the evidence body (`comment.md`, and `comment.txt` where this skill uses it):
48
+
49
+ - It contains a `## Not established` heading. That heading is **never omitted and never blank** — when nothing is outstanding it still renders `None outstanding — reviewed`; otherwise it names, in plain operator language, what the verification did not prove.
50
+ - The accompanying verdict carries `not_established_reviewed: true` (the list may be empty; the flag may never be omitted).
51
+ - It contains a `## Artifact identity` heading carrying **values, not placeholders** — the repository, the `head_sha` the verification observed, the `environment`, and per artifact its `sha256` digest and `captured_at`. **Refuse to post** a body whose identity heading is absent or unpopulated, or whose recorded `artifact_head_sha` disagrees with the verdict's `artifact.head_sha` — report the evidence id and **both SHAs**. Definition: the `claim-evidence-mapping` rule.
52
+
53
+ If either is missing, **refuse to post**: stop and report the missing Not-established review to the caller instead of publishing. Composing the body is `lisa-tracker-evidence`'s job (see its UI Evidence Checklist); this skill only refuses to publish one that omits the section. The section is defined by the `claim-evidence-mapping` rule and generalizes `lisa-improve-harness`'s required, never-empty `Known limits` field.
54
+
45
55
  ## Usage
46
56
 
47
57
  ```bash
@@ -41,6 +41,16 @@ The caller must produce:
41
41
 
42
42
  If any of these are missing, stop and report.
43
43
 
44
+ ## Comment-body preflight (required)
45
+
46
+ Before posting or updating anything, check the evidence body (`comment.md`, and `comment.txt` where this skill uses it):
47
+
48
+ - It contains a `## Not established` heading. That heading is **never omitted and never blank** — when nothing is outstanding it still renders `None outstanding — reviewed`; otherwise it names, in plain operator language, what the verification did not prove.
49
+ - The accompanying verdict carries `not_established_reviewed: true` (the list may be empty; the flag may never be omitted).
50
+ - It contains a `## Artifact identity` heading carrying **values, not placeholders** — the repository, the `head_sha` the verification observed, the `environment`, and per artifact its `sha256` digest and `captured_at`. **Refuse to post** a body whose identity heading is absent or unpopulated, or whose recorded `artifact_head_sha` disagrees with the verdict's `artifact.head_sha` — report the evidence id and **both SHAs**. Definition: the `claim-evidence-mapping` rule.
51
+
52
+ If either is missing, **refuse to post**: stop and report the missing Not-established review to the caller instead of publishing. Composing the body is `lisa-tracker-evidence`'s job (see its UI Evidence Checklist); this skill only refuses to publish one that omits the section. The section is defined by the `claim-evidence-mapping` rule and generalizes `lisa-improve-harness`'s required, never-empty `Known limits` field.
53
+
44
54
  ## Phase 1 — Resolve Linear Issue
45
55
 
46
56
  1. Parse the identifier from `$ARGUMENTS`.
@@ -28,6 +28,10 @@ See the `config-resolution` rule for configuration and dispatch table.
28
28
  - Never post evidence to a different ticket than the one named — `$ARGUMENTS` is the source of truth.
29
29
  - Never invent a verify-specific usage footer. Evidence artifact usage must flow through `lisa-usage-accounting`, preserve the canonical `## Lisa Usage` section, and surface `source: unavailable` explicitly when the runtime cannot provide trustworthy numbers.
30
30
  - **Evidence-manifest gate (leaf work units).** Before dispatching to a vendor skill that transitions the ticket, confirm `EVIDENCE_DIR` contains a non-empty artifact **of the declared type** for every typed `[EVIDENCE: <artifact-type>: <name>]` marker declared in the ticket's Validation Journey — a `screenshot` marker needs an actual image, an `http-transcript` marker needs the request + response text, a `perf-trace` marker needs measured numbers; a prose claim satisfies nothing. If any declared marker has no captured artifact, an empty one, or one whose content/extension does not match its declared type, stop and report the offending markers by name instead of posting — a leaf work unit (Bug / Task / Sub-task / Improvement) may not advance to its review/Done state with an unsatisfied manifest (see the "Per-Work-Unit Evidence Contract" in the `verification` rule). Epics / Stories / Spikes, and leaf units without a Validation Journey, are exempt.
31
+ - **Claim↔boundary binding (S14 upgrade).** Satisfying the manifest by *type* is not enough: each `[EVIDENCE: <artifact-type>: <name>]` marker is also cited for a **claim**, and the marker's artifact type must **reach the boundary that claim declares** per the `claim-evidence-mapping` rule's taxonomy. A `browser` claim (user-visible UI behavior) is reached by `screenshot` / `recording` and never by a unit `test-run-log`; an `http-api` claim needs an `http-transcript`; a `deploy-health` claim needs a `deploy-log` and no pre-deploy artifact. On a mismatch, report the offending marker **by name** together with the claim, its boundary, and the **required evidence kinds** for that boundary — e.g. `[EVIDENCE: test-run-log: unit-suite] cited for claim AC-2 [boundary browser] — required evidence kinds: screenshot, recording`. Swapping in a marker of a reaching kind with a real captured artifact satisfies the gate. **Advisory-first:** until `verification.gate.enforceBoundaries` is `true` in `.lisa.config.json` (the same ratchet flag the Stop-hook gate reads, default `false`), a boundary mismatch is reported to the operator but does not block the post; once ratcheted on it refuses the post exactly like a missing or wrong-type artifact. Missing/empty/wrong-type artifacts keep refusing the post regardless of the flag.
32
+ - **The "Not established" section is required.** Before posting, confirm `evidence/comment.md` (and `comment.txt` where the vendor uses it) contains a `## Not established` heading and that the verdict carries `not_established_reviewed: true`. The heading is **never omitted and never blank**: with nothing outstanding it still renders `None outstanding — reviewed`; otherwise it lists, in plain operator language, each thing the verification did not prove — boundaries not exercised, environments not tested, behavior consciously out of scope. A comment with no such heading, a heading with nothing under it, or a verdict whose `not_established_reviewed` flag is absent is **refused**: stop and report the missing Not-established review instead of posting. (The list may be empty; the flag may never be omitted.) Definition and exemplars live in the `claim-evidence-mapping` rule; this generalizes `lisa-improve-harness`'s required, never-empty `Known limits` field.
33
+ - **Artifact identity is required, with values (S14 extension).** Before posting, confirm the comment body carries a `## Artifact identity` heading populated with **values, not placeholders**: the `repository`, the `head_sha` the verification observed, the `environment`, and for each committed artifact its `sha256` digest and `captured_at`. Then check both identity failures. **`artifact_mismatch`** — an evidence entry whose recorded `artifact_head_sha` differs from the verdict's `artifact.head_sha`: refuse the post and report the offending evidence id together with **both SHAs** (the one the artifact was captured at and the one the verdict claims), e.g. `EV-2 captured at 4f1c9ab but artifact.head_sha is 9de0c31`. **`evidence_digest_mismatch`** — recompute the `sha256` of each committed evidence file on read; if the bytes no longer match the recorded digest (or the file is absent), stop and report the evidence id by name. Never summarize either as "verification failed" — name the field, the ids, and both SHAs so a non-engineer can read it at the gate. **Advisory-first:** until `verification.gate.enforceBoundaries` is `true` in `.lisa.config.json`, an identity failure is reported to the operator but does not block the post; once ratcheted on it refuses the post exactly like a missing artifact. Definition: the `claim-evidence-mapping` rule.
34
+ - **Merge-race reconciliation (at completion, not at post).** The evidence's pinned `head_sha` is reconciled against the merged head using the ancestry + deploy-run definition of "what shipped" that `lisa-drive-pr-to-merge` already owns — cite that skill, never re-implement or restate its checks here. Pre-merge evidence counts for the merge commit only when its head is a parent of the merge, ancestry alone never justifies reporting shipped, and on a mismatch verification re-runs against the merged head before completion is declared.
31
35
  - **Evidence references are not manifest entries.** Extract obligations using the exact `[EVIDENCE:` prefix (and the legacy local `[SCREENSHOT:` form). Exclude both the canonical `[EVIDENCE-REF: <work-item-ref> | <artifact-type>: <kebab-case-name>]` and the Lisa 2.223.0 legacy alias `[EVIDENCE-REF: <tracker-ref>: <artifact-type>: <kebab-case-name>]` from artifact lookup, missing-artifact reporting, and duplicate-name checks: either belongs to another work item and cannot satisfy this item's S14 gate. A runtime-changing leaf with references but no local claiming marker must be rejected before dispatch.
32
36
 
33
37
  ## UI Evidence Checklist (when work is UI-visible)
@@ -47,8 +51,11 @@ The checklist is tracker-agnostic — the same shape works on JIRA, GitHub Issue
47
51
  5. **"What this shows" section.** Tailor to ticket type:
48
52
  - **Bug repro:** state plainly whether the bug reproduces or not, and the most likely 1–2 reasons their retest still failed (different env, native app vs. web, stuck backend row, etc.).
49
53
  - **Feature/UX completion:** state plainly which acceptance criteria each screenshot covers, and call out any deferred or out-of-scope surface explicitly so QA/PM doesn't have to infer.
50
- 6. **"What would help me confirm" (bug) / "How to QA" (feature) section.** Concrete actionable retest steps with the exact selection criteria (e.g., "pick a record whose Status column shows `—`, not `Processing` or `Pending Review`").
51
- 7. **Explicit invitation to be corrected.** End with a line like *"If any of the steps I listed are different from what you expected / actually did, please tell me explicitly which step I got wrong."* Non-optional small differences (which record, which device, exact tap order, expected behavior) change everything, and naming the door open short-circuits ticket bounce-loops.
52
- 8. **Workflow transition** is the vendor skill's job, not yoursit'll move the ticket per the configured tracker (JIRA: Reassign to reporter for bug repro / move to the configured review status when one exists, otherwise leave it in `claimed`; GitHub: direct `claimed` configured `done` after a successful build; Linear: equivalent state). You don't transition manually.
54
+ 6. **"Artifact identity" and "Not established" sections (both required).** Right after "What this shows":
55
+ - `## Artifact identity` what the evidence was collected against, as **values, not placeholders**: the repository, the `head_sha` the verification observed, the `environment`, and per artifact its `sha256` digest and `captured_at`. A heading rendered with `<unknown>` in place of a SHA is not identity.
56
+ - `## Not established` — **never omitted, never blank.** List in plain language what this verification did *not* prove: boundaries not exercised (e.g. "the persisted order row was never queried"), environments not tested (e.g. "checked on desktop Chrome only not mobile Safari"), and behavior consciously out of scope (e.g. "refunds were not touched or tested"). Name the specific thing, not a category. With nothing outstanding, the heading still renders a single line: `None outstanding reviewed`. Never state a quality check as proof of behavior "unit tests pass" belongs here as a limit, not above as evidence.
57
+ 7. **"What would help me confirm" (bug) / "How to QA" (feature) section.** Concrete actionable retest steps with the exact selection criteria (e.g., "pick a record whose Status column shows `—`, not `Processing` or `Pending Review`").
58
+ 8. **Explicit invitation to be corrected.** End with a line like *"If any of the steps I listed are different from what you expected / actually did, please tell me explicitly which step I got wrong."* Non-optional — small differences (which record, which device, exact tap order, expected behavior) change everything, and naming the door open short-circuits ticket bounce-loops.
59
+ 9. **Workflow transition** is the vendor skill's job, not yours — it'll move the ticket per the configured tracker (JIRA: Reassign to reporter for bug repro / move to the configured review status when one exists, otherwise leave it in `claimed`; GitHub: direct `claimed` → configured `done` after a successful build; Linear: equivalent state). You don't transition manually.
53
60
 
54
61
  **Why this format:** It (a) gives the reporter a frame-by-frame they can compare against, (b) avoids the JIRA image-collapse failure mode while still working everywhere else, (c) names the most plausible discrepancies up front so the loop short-circuits, (d) explicitly opens the door to being corrected so tickets don't bounce on assumed alignment. The same mechanics that resolve a stuck bug ticket also give QA an unambiguous handoff for a freshly-built feature.