fdeops 5.1.0 → 5.1.1

This diff represents the content of publicly available package versions that have been released to one of the supported registries. The information contained in this diff is provided for informational purposes only and reflects changes between package versions as they appear in their respective public registries.
Files changed (46) hide show
  1. package/README.md +9 -3
  2. package/mcp/fdeops-ingest/package.json +1 -1
  3. package/package.json +1 -1
  4. package/plugin.json +1 -1
  5. package/skills/brief/.fde-generated.json +1 -1
  6. package/skills/brief/references/land.md +28 -28
  7. package/skills/build/.fde-generated.json +2 -2
  8. package/skills/build/references/build.md +1 -1
  9. package/skills/build/references/integrate.md +12 -2
  10. package/skills/debug/.fde-generated.json +2 -2
  11. package/skills/debug/references/build.md +1 -1
  12. package/skills/debug/references/integrate.md +12 -2
  13. package/skills/earn-trust/.fde-generated.json +1 -1
  14. package/skills/earn-trust/references/earn-trust.md +27 -46
  15. package/skills/evaluate/.fde-generated.json +2 -2
  16. package/skills/evaluate/references/build.md +1 -1
  17. package/skills/evaluate/references/integrate.md +12 -2
  18. package/skills/fde/SKILL.md +15 -13
  19. package/skills/fde/references/build.md +1 -1
  20. package/skills/fde/references/earn-trust.md +27 -46
  21. package/skills/fde/references/hold-scope.md +25 -24
  22. package/skills/fde/references/integrate.md +12 -2
  23. package/skills/fde/references/land.md +28 -28
  24. package/skills/fde/references/rescue.md +18 -18
  25. package/skills/fde/references/who-decides.md +29 -57
  26. package/skills/integrate/.fde-generated.json +2 -2
  27. package/skills/integrate/references/build.md +1 -1
  28. package/skills/integrate/references/integrate.md +12 -2
  29. package/skills/poc/.fde-generated.json +2 -2
  30. package/skills/poc/references/build.md +1 -1
  31. package/skills/poc/references/integrate.md +12 -2
  32. package/skills/qa/.fde-generated.json +2 -2
  33. package/skills/qa/references/build.md +1 -1
  34. package/skills/qa/references/integrate.md +12 -2
  35. package/skills/rescue/.fde-generated.json +1 -1
  36. package/skills/rescue/references/rescue.md +18 -18
  37. package/skills/review/.fde-generated.json +2 -2
  38. package/skills/review/references/build.md +1 -1
  39. package/skills/review/references/integrate.md +12 -2
  40. package/skills/scope/.fde-generated.json +1 -1
  41. package/skills/scope/references/hold-scope.md +25 -24
  42. package/skills/ship/.fde-generated.json +2 -2
  43. package/skills/ship/references/build.md +1 -1
  44. package/skills/ship/references/integrate.md +12 -2
  45. package/skills/who-decides/.fde-generated.json +1 -1
  46. package/skills/who-decides/references/who-decides.md +29 -57
package/README.md CHANGED
@@ -34,7 +34,7 @@ Help me prepare for the first meeting. Here is the brief: …
34
34
 
35
35
  `fde` asks for the information needed next and uses the relevant instructions. As work progresses, it can help you investigate delays, compare approaches, implement a change, test it, and prepare a customer update. You do not have to choose a skill at each step.
36
36
 
37
- Naming the customer starts a local project record at `~/fde-engagements/client01/.fde/`. It keeps the brief, decisions, evidence and next actions together. Review proposed agreements and corrections before saving them. [How customer records work](#keep-a-customer-record).
37
+ For an ongoing project, naming the customer starts a local record at `~/fde-engagements/client01/.fde/`. It keeps the brief, decisions, evidence and next actions together. Review proposed agreements and corrections before saving them. [How customer records work](#keep-a-customer-record).
38
38
 
39
39
  ### Use just one task
40
40
 
@@ -77,7 +77,7 @@ Use the skill that matches the work in front of you. All skills are individually
77
77
 
78
78
  These examples are starting points, not a required sequence. The catalog also covers inherited projects, business cases, incidents, recovery, source connections and multiple customer records. `fde` uses the same instructions and loads additional detail only when needed.
79
79
 
80
- **Which mode should I choose?** Use an individual skill for a specific task. Use `fde` when you want help across several tasks and a continuing customer record. Installing a skill makes it available independently; it does not remove its real input requirements. For example, `dashboard` needs records to display, while `debrief` can review pasted notes without saving anything.
80
+ **Which mode should I choose?** Use an individual skill for a specific task. Use `fde` when you want help selecting the next task or maintaining a continuing customer record. You can also ask it for a one-off result without creating records. Installing a skill makes it available independently; it does not remove its real input requirements. For example, `dashboard` needs records to display, while `debrief` can review pasted notes without saving anything.
81
81
 
82
82
  <a name="how-skills-work"></a>
83
83
 
@@ -125,7 +125,13 @@ Open a customer's record and copy an action into your agent to continue. Run the
125
125
 
126
126
  ## Fit it to the project
127
127
 
128
- For a small task, use the supplied notes or code and return the result. For an ongoing project, use the customer record. For enterprise work, include the relevant teams, access rules, release checks and operating responsibilities in the plan.
128
+ Start with the work in front of you:
129
+
130
+ | Project | How to use FDEOps |
131
+ |---|---|
132
+ | Small fix or analysis | Give a task skill the relevant notes or code. Get the result and its verification; no customer record is required. |
133
+ | Customer integration or ongoing delivery | Use `fde` to connect discovery, implementation, tests and updates in one continuing record. |
134
+ | Enterprise engagement | Work within the customer’s existing access, change-control and operating processes. Record who can approve each decision and what evidence they require. |
129
135
 
130
136
  The pack includes implementation, integration, debugging and QA instructions. It uses the repository's existing tools. It does not supply customer credentials, infrastructure, specialist approvals or production authority.
131
137
 
@@ -1,6 +1,6 @@
1
1
  {
2
2
  "name": "fdeops-ingest-mcp",
3
- "version": "5.1.0",
3
+ "version": "5.1.1",
4
4
  "private": true,
5
5
  "description": "Thin stdio MCP sink for FDEOps ingest (stage \u2192 propose \u2192 apply). Zero runtime dependencies.",
6
6
  "bin": {
package/package.json CHANGED
@@ -1,6 +1,6 @@
1
1
  {
2
2
  "name": "fdeops",
3
- "version": "5.1.0",
3
+ "version": "5.1.1",
4
4
  "description": "Forward deployed engineering skills for AI coding agents. Use focused task skills or @fde for discovery, implementation, verification and handoff, with local engagement records.",
5
5
  "bin": {
6
6
  "fdeops": "bin/install.js",
package/plugin.json CHANGED
@@ -1,7 +1,7 @@
1
1
  {
2
2
  "$schema": "https://agent-plugins.org/schemas/1.0.0/plugin.schema.json",
3
3
  "name": "fdeops",
4
- "version": "5.1.0",
4
+ "version": "5.1.1",
5
5
  "description": "Forward deployed engineering skills for AI coding agents. Use focused task skills or @fde for discovery, implementation, verification and handoff, with local engagement records.",
6
6
  "author": {
7
7
  "name": "Subash Natarajan",
@@ -3,7 +3,7 @@
3
3
  "version": 1,
4
4
  "files": {
5
5
  "SKILL.md": "6ccada100587e370639f468328795839c15a063b7ada4ed2f31794ba9e6c1c4c",
6
- "references/land.md": "be633eefbfdb202638f4a04ba2f1320ef69a5af33f7aa685a3c7d95aa5d5c6f9",
6
+ "references/land.md": "7a82ac999d23b51b602cd815777a853c4879ca8be40dc6c84c66f154d1f77951",
7
7
  "references/task-context.md": "73eea2d7f164fac3226599e5be26ae4e79dcf69e0d24428dd12d623861410490"
8
8
  }
9
9
  }
@@ -2,7 +2,7 @@
2
2
 
3
3
  **Enter when:** new customer, first meeting, just got the brief, nothing started yet.
4
4
 
5
- **Read first:** `context.md` if it exists, then the supplied brief. Once the engagement type and AI/access policy are known, inspect the supplied repo/docs relevant to the ask before asking questions they can answer. This is a bounded evidence check, not a full discovery scan.
5
+ **Read first:** apply [task context](task-context.md), then permitted `context.md` evidence if it exists and the supplied brief. Once the engagement type and AI/access policy are known, inspect the supplied repo/docs relevant to the ask before asking questions they can answer. This is a bounded evidence check, not a full discovery scan.
6
6
 
7
7
  ## Validation gate (confirm understanding, clarify where it elevates)
8
8
 
@@ -26,27 +26,27 @@ Format - one question at a time, with a guess the FDE can correct:
26
26
 
27
27
  ```
28
28
  READ: <one sentence - what you think they actually need>
29
- CONFIDENCE: ~NN% - missing: <what still blocks a safe start>
29
+ MISSING: <fact or authority that changes the next action>
30
30
  Q: <one focused question>
31
- GUESS: <your best answer, so they can push back fast>
31
+ POSSIBLE READ: <clearly labeled interpretation, if useful; never guessed authority>
32
32
  ```
33
33
 
34
- Wait for the reaction before the next question. Stop when confidence is high enough to write `success.md` without inventing names, or when the FDE says move on. Every answer that is still unknown stays `unknown - ask:` in the artifact - never fill the gap with a plausible stakeholder.
34
+ Wait for the reaction before the next question. Stop when the next authorized action is clear, or when the FDE says move on; unanswered material gaps remain visible. Every answer that is still unknown stays `unknown - ask:` in the artifact - never fill the gap with a plausible stakeholder.
35
35
 
36
36
  ## Method - part 1: interrogate the brief (you do this work)
37
37
 
38
38
  Read the brief the FDE gives you. Separate **observed** (source/path and date), **reported** (who said it), and **hypothesis** (how to test it). A requested solution such as “build an agent” is not evidence of the cause. Ask only about gaps that change scope, access, acceptance, or the next investigation. What is **not** in the brief matters as much as what is. Produce the gap list yourself:
39
39
 
40
- - **No named decision-maker** → the FDE will spend two weeks building for someone who can't say yes. Flag it.
41
- - **"Straightforward cleanup" on an 8-year-old system** → the previous attempt is still visible in git history as a revert. Flag it.
42
- - **Very tight timeline** → someone already promised the outcome before hiring the FDE. Flag it.
43
- - **No out-of-scope section** → scope creep is pre-authorized. Flag it.
40
+ - **No named decision-maker** → identify who or what can accept the outcome and the source of that authority; keep it unknown until established.
41
+ - **"Straightforward cleanup" on an 8-year-old system** → inspect permitted relevant history and tests for prior attempts and constraints; age alone proves neither complexity nor a previous failure.
42
+ - **Very tight timeline** → establish the deadline, its source and which commitments are actually agreed.
43
+ - **No out-of-scope section** → clarify material boundaries against the existing agreement. An omission does not authorize additional work.
44
44
 
45
45
  Write these into `brief.md` as **questions to answer**, not problems - they're what the FDE is walking in to resolve.
46
46
 
47
47
  Pre-arrival checks to run through with the FDE:
48
- - Access confirmed? Repo, environment, docs. Waiting for access on day two burns trust.
49
- - Has someone tried this before? Find out why it failed before assuming this approach is different.
48
+ - Access confirmed for the next task? Repo, environment and docs may have different permissions; identify gaps before dependent work.
49
+ - Has someone tried this before? Establish what happened and what evidence remains; do not assume the attempt failed.
50
50
  - Other vendors/teams in scope? Then the FDE is not the only one in the room, even when alone in the meeting.
51
51
  - Tech stack recon: job postings, GitHub org - know the stack before they say it.
52
52
 
@@ -59,21 +59,21 @@ Intent: coach the FDE's first *customer* conversation - what keeps the sponsor u
59
59
  - "Who loses credibility if this goes wrong?"
60
60
  - "If nothing changes over the agreed timeframe, what happens, and who bears it?" Record the consequence and its source in `brief.md`; distinguish reported impact from measured cost. Unknown cost stays unknown, not an invented ROI.
61
61
 
62
- Let silence sit. If their fear doesn't match the written brief, the brief is wrong - say so plainly, log it.
62
+ Allow time for an answer. If a stated concern differs from the brief, record the difference and clarify whether it changes the agreed outcome; neither statement automatically supersedes the other.
63
63
 
64
64
  **Listen for, and capture as you hear it:**
65
- - **The real decision-maker** - whoever others mention most, especially if not yet met. That's who judges the work.
66
- - **The previous attempt** - "we tried something similar last year" is the most important sentence in the first meeting. Who was involved? Still there and protective, or gone because of it?
67
- - **The passed-over internal team** - they know exactly what's wrong, and they resent the FDE's presence. Find them before the first standup, ask what they tried, use their language in every meeting. Make them look right and they protect you; ignore them and they wait for the mistake.
68
- - **The sacred thing** - "Is there anything in this environment I should treat as untouchable?" The hesitation before the answer is the answer.
65
+ - **Decision rights** - who can approve scope, accept the result and authorize release, as relevant. A frequently mentioned person may be influential; confirm their actual authority and scope.
66
+ - **The previous attempt** - "we tried something similar last year" identifies evidence to investigate. Who was involved, what happened, and which constraints still apply? Do not infer why someone left.
67
+ - **The existing internal team** - ask what they tried, what they know and what they expect to own. Use established terminology and credit their work. Do not assume resentment, displacement or complete knowledge of the problem.
68
+ - **The sacred thing** - "Is there anything in this environment I should treat as untouchable?" Capture the stated boundary and applicable policy; hesitation alone does not identify a restriction.
69
69
  - **Exception path (operating map seed)** - "When the happy path breaks this week, what do people actually do - who do they call, what spreadsheet opens, what do they skip?" Capture the break → workaround → who owns it. Do not build a full map on day 1; seed rows later in `terrain.md` → `## Operating map (exception-led)` during discover. Unknowns stay `unknown - ask:`.
70
70
  - **AI posture and policy** - tools already in use (sanctioned or shadow), and: "Does your organisation have a policy on AI-generated code? Are there decisions where you would not be comfortable with AI involvement?"
71
71
  - **Future operator** - "Who will run this after we leave, and have they agreed?" Record the proposed operator and unresolved ownership in `success.md`, separately from the signer. A sponsor naming a team is not that team accepting responsibility; verify with the operator during discover.
72
72
  - **Boundaries in multi-vendor rooms** - who owns what surface, who signs off before a change crosses it.
73
73
 
74
- ## The day 1 deliverable
74
+ ## An early deliverable
75
75
 
76
- After `success.md` names a signer (or the FDE explicitly overrides with `unknown - ask:` still visible), ship one visible thing before the end of day 1: a small bug fix, a cleanup the team has stepped over, a dashboard tweak, a config improvement. Not because it matters technically - because it proves you can ship in their environment without breaking things. The first deploy sets the trust trajectory for the entire engagement. A day-1 deliverable earns more credibility than a week-3 architecture deck. Skip it until the land gate is met.
76
+ Choose an early useful result within confirmed scope: a verified small fix, a permitted diagnostic, or a concise map of an unresolved problem. Reuse the existing outcome and authority for routine work. A first-day deadline does not grant deployment permission or waive verification; use `ship` for a release. If a missing signer blocks a consequential decision, keep it visible and continue independent preparation.
77
77
 
78
78
  ## Artifact (write as the conversation is debriefed)
79
79
 
@@ -81,7 +81,7 @@ After `success.md` names a signer (or the FDE explicitly overrides with `unknown
81
81
 
82
82
  **`success.md`** - what done looks like, **primary value bucket** (`cost-save` | `risk-mitigation` | `revenue-uplift`), baseline → target, who actually signs off, what is explicitly out of scope. Record agreement only with its source and scope; otherwise label the target proposed. For each baseline, record source, date/window, environment, and sample size when relevant. An operator recollection is reported, not measured. If no baseline exists, name the measurement owner and cheapest way to obtain it; do not manufacture a number.
83
83
 
84
- For every target number, run the **gaming check** before it is written down: *how could this metric hit its target without the customer being any better off?* There is always an answer, and the answer is what the org will drift toward under pressure. Write the guard next to the metric:
84
+ For every target number, run the **gaming check** before it is written down: *how could this metric hit its target without the customer being any better off?* Identify plausible failure modes without predicting that the customer will exploit them. Write a relevant guard next to the metric:
85
85
 
86
86
  ```markdown
87
87
  | Metric | Baseline → target | Gamed by | Guard |
@@ -89,19 +89,19 @@ For every target number, run the **gaming check** before it is written down: *ho
89
89
  | reconciliation alert latency | 4h → 15min | alerting on everything, so nobody reads them | alerts acked by a named owner, ≤2/week |
90
90
  ```
91
91
 
92
- A metric with no gaming check is a metric the FDE will be held to and cannot defend. If the customer resists the guard, that is the real conversation - they are attached to the number, not the outcome.
92
+ If a proposed guard is disputed, capture the stated reason and assess its cost and effect on the outcome. Do not infer that the customer values the number over the result.
93
93
 
94
94
  **`stakeholders.md`**:
95
95
  ```markdown
96
96
  | Who | Role | Signal | Notes |
97
97
  |-----|------|--------|-------|
98
- | <name> | sponsor / champion / resistor / veto / passed-over | green/amber/red | <evidence, day> |
98
+ | <name> | <observed participation role; authority recorded separately> | green/amber/red | <evidence, day> |
99
99
  ```
100
100
  If `stakeholders.md` already has a `## Signal history` section (it does from the template), **never delete or overwrite it** when you rewrite this file - it holds the dated `[signal:...]` tokens `fde log contact --signal` and `fde debrief` write, and `fde status`/`fde receipts`/the dashboard read only from that section. Edit the table above it freely; keep the section below intact.
101
101
 
102
102
  **`trust-profile.md`** - sacred data (`<private>` tagged), fears heard, AI policy, approval chain. Sensitive: skip for status reads; use CLI/redacted surfaces; never paste raw `<private>` into prompts or subagents.
103
103
 
104
- **`assumptions.md`** - seed every unverified claim from the brief (and the day-1 hypothesis) as rows with Kind `UNKNOWN` (or `CONVENTION` if they said "we always"), blast radius CRITICAL / LOAD-BEARING / CONVENIENCE, and status `OPEN`. Do not wait for test-assumptions - land makes the register exist. Example:
104
+ **`assumptions.md`** - seed consequential unverified claims from the brief (and the initial hypothesis) as rows with Kind `UNKNOWN` (or `CONVENTION` if they said "we always"), blast radius CRITICAL / LOAD-BEARING / CONVENIENCE, and status `OPEN`. Do not wait for test-assumptions - land makes the register exist. Example:
105
105
 
106
106
  ```markdown
107
107
  | # | Assumption | Kind | Blast radius | How we test | Status | Evidence |
@@ -113,7 +113,7 @@ One falsifiable hypothesis about the real problem also goes at the bottom of `br
113
113
 
114
114
  ## Checkpoint
115
115
 
116
- One page back to the FDE: success + value bucket + sign-off owner, out-of-scope boundary, sacred data, stakeholder map with veto power, AI posture, the hypothesis, the top CRITICAL assumptions still OPEN, and any exception-path seeds heard (break → workaround → owner) for discover to map into `terrain.md`. If it doesn't fit one page, the engagement isn't understood yet.
116
+ One page back to the FDE: success + value bucket + sign-off owner, out-of-scope boundary, sacred data, stakeholder map with veto power, AI posture, the hypothesis, the top CRITICAL assumptions still OPEN, and any exception-path seeds heard (break → workaround → owner) for discover to map into `terrain.md`. Keep the summary short and link necessary detail; a complex engagement may need supporting evidence.
117
117
 
118
118
  If remote: agree how progress and blockers will be shared; use a short call when asynchronous context is insufficient.
119
119
 
@@ -121,16 +121,16 @@ If remote: agree how progress and blockers will be shared; use a short call when
121
121
 
122
122
  Kickoff at Acme payments. Priya (VP Eng) sponsors; the brief says "add monitoring to the reconciliation service."
123
123
 
124
- Asking what happens the week after a perfect delivery gets: "I stop hearing about it from finance." That is the real success statement - not monitoring. The previous attempt surfaces too: the platform team built alerting last year, it was turned off. Raj, who built it, is still there and was not in the kickoff - the passed-over team, found on day 1 rather than at the first standup.
124
+ Asking what happens the week after a perfect delivery gets: "I stop hearing about it from finance." That suggests a concern to clarify alongside the monitoring request. The previous attempt surfaces too: the platform team built alerting last year, it was turned off. Raj, who built it, is still there and was not in the kickoff. Ask for his account of the earlier attempt; his absence does not explain his views.
125
125
 
126
- What gets written: `success.md` with bucket `risk-mitigation`, `reconciliation failures reach a named owner within 15 min (baseline: 4h, found by finance)`, gaming check `alerting on everything so nobody reads them` → guard `≤2 alerts/week, acked by name`, sign-off Priya. `brief.md` carries the gap list and the hypothesis: *the job is not unmonitored, it is unowned*. `assumptions.md` seeds `"finance would act on an alert" - CRITICAL - OPEN - (stated, unverified)`. `trust-profile.md` records the sacred thing Priya hesitated before naming.
126
+ In this example Priya reports a four-hour baseline and proposes the following target; her acceptance authority still needs its source. What gets written: `success.md` with proposed bucket `risk-mitigation`, `reconciliation failures reach a named owner within 15 min (baseline: 4h, found by finance)`, gaming check `alerting on everything so nobody reads them` → guard `≤2 alerts/week, acked by name`, proposed sign-off Priya until confirmed. `brief.md` carries the gap list and the hypothesis: *the job is not unmonitored, it is unowned*. `assumptions.md` seeds `"finance would act on an alert" - CRITICAL - OPEN - (stated, unverified)`. `trust-profile.md` records the boundary Priya explicitly names, through the permitted privacy-safe workflow.
127
127
 
128
- Day-1 deliverable: fix the log line that swallows the job's exit code. Small, visible, in their environment.
128
+ Early deliverable: verify and fix the log line that swallows the job's exit code within the existing scope. Deployment remains subject to the established release authority and checks.
129
129
 
130
130
  ## Principles
131
131
 
132
- - Never start technical work before `success.md` exists.
132
+ - Establish the outcome and authority needed for the next action; missing record files do not block useful standalone work.
133
133
  - Sacred data tagged `<private>` stays out of model context: use CLI/redacted reads; never paste raw private blocks.
134
- - The brief is a hypothesis; discover confirms it. Seed `assumptions.md` on day one.
135
- - The passed-over internal team is the best source of truth, not an obstacle.
134
+ - Treat unverified parts of the brief as hypotheses; discovery may support or overturn them. Record consequential assumptions.
135
+ - Learn from the existing team and verify consequential claims without guessing motives.
136
136
  - If the customer cannot define success, that is the first problem to solve.
@@ -3,10 +3,10 @@
3
3
  "version": 1,
4
4
  "files": {
5
5
  "SKILL.md": "19f32ecae72b96ee19cb87658ddd998cce96ec2e3960255d13fcc2906c6d0b6b",
6
- "references/build.md": "3dfeed619eeb1c8401f5cdf65e6f803fb209c70cb464dac4e60a1c890fd3a6f7",
6
+ "references/build.md": "6fa07b15b682119de53c38950f94942a58e4680ab58b24319d77c54a27cd9697",
7
7
  "references/debug.md": "c3bb344d38cc3552cb4e230c601a9be3fe173af2b2efb89aab6b7b04339f24f4",
8
8
  "references/eval-pack.md": "0590b85d3cae0903c6b1274540c92eaa2a4373047e8a0548d6942516ef0bb9e1",
9
- "references/integrate.md": "107a50bddf6cb0ba7f2bc006dfe9851800c6868e43f785a33ae4b737aeb74c95",
9
+ "references/integrate.md": "d0de35a783902ca8b4762e3a42a14f467766d56a928c3a5cf11adac2a6ba90ba",
10
10
  "references/qa.md": "d8f58e6d36436469a58aeb1107037f3e27fa81ff5b82d0e4df3c23eeadaf683c",
11
11
  "references/review.md": "63a007f78288089cc84cccc72647e8ce6721b7efa0f4f8d6774c0f0af594749d",
12
12
  "references/ship.md": "8cdcb2d4d6eb57e0adf3f1996bc02ae66920852ca304d2afd778fa483b7e969a",
@@ -6,7 +6,7 @@ Use the permitted context and authority in [task context](task-context.md). This
6
6
 
7
7
  ## Method
8
8
 
9
- 1. Identify the repository, its instructions, working tree, relevant callers, and test commands. Inspect examples before creating abstractions. Preserve unrelated edits and state which dependencies or interfaces the change touches.
9
+ 1. Identify the repository, its instructions, working tree, relevant callers, and test commands. Inspect examples before creating abstractions. Preserve unrelated edits and state which dependencies or interfaces the change touches. Before changing an untested legacy path, capture the undocumented behavior callers depend on with targeted characterization checks; distinguish behavior to preserve from the intended change.
10
10
  2. State the observable outcome, constraints, and acceptance checks. Reuse agreed criteria for routine fixes. If a consequential product choice is unresolved, surface that choice while continuing independent investigation; do not invent acceptance.
11
11
  3. Choose the smallest coherent path that demonstrates the outcome through the real entry point. Include the necessary storage, error handling, and interface behavior in that slice. Name the failure that stops expansion and the recovery path for stateful changes.
12
12
  4. Implement using the repository's tools and conventions. Search for existing services, fixtures, and validation before adding alternatives. Keep cleanup limited to what makes the changed path understandable; do not expand scope to repair unrelated code.
@@ -6,13 +6,23 @@ Start from [task context](task-context.md). Permitted supplied context is enough
6
6
 
7
7
  ## Method
8
8
 
9
- 1. Map producer, consumer, owner, direction, and side effects. Inspect the actual installed version and local implementation; verify uncertain behavior against current official documentation. Identify the relevant schema, authentication scopes, network boundary, and permitted test environment.
9
+ 1. Map producer, consumer, owner, direction, and side effects. Inspect the actual installed version and local implementation; verify uncertain behavior against current official documentation. Identify the relevant schema, authentication scopes, network boundary, and permitted test environment. Before changing an untested existing boundary, characterize the mappings, ordering or other observable behavior its callers depend on.
10
10
  2. Write the acceptance example: an input at the real boundary and the observable downstream result. Include a rejection or failure example. Separate configuration validity, successful authentication, transport connectivity, contract compatibility, and end-to-end behavior; none proves the next.
11
11
  3. Inspect credentials by presence and required scope without printing values. Use existing secret storage. Check data classification and retention before moving data; never pass raw `<private>` blocks into a model. Prefer sanitized or synthetic cases approved for the target environment.
12
- 4. Implement the narrow adapter using native repository patterns. Validate external inputs and model outputs, bound timeouts and retries, preserve error context without leaking payloads, and handle cancellation. For writes, establish idempotency or duplicate detection before retries; for events, check ordering, replay, and poison messages as applicable.
12
+ 4. Implement the narrow adapter using native repository patterns. Validate external inputs and model outputs, bound timeouts and retries, preserve error codes and failure phase without leaking payloads, and handle cancellation. Keep explicit authentication or permission rejections distinguishable from transport uncertainty. For writes, establish idempotency or duplicate detection before retries and apply the uncertain-write rules below when outcomes can be ambiguous; for events, check ordering, replay, and poison messages as applicable.
13
13
  5. Exercise a permitted success case and relevant failures: denied access, malformed data, rate limit, timeout, duplicate delivery, or partial completion. Trace correlation IDs or safe evidence across both sides. A mock proves client behavior only; if live access is unavailable, report that gap instead of claiming an integration works.
14
14
  6. Check cleanup and recovery for test side effects. Use [verification](verification.md) for receipts and [review](review.md) for security or data-contract changes. Route deployment through [ship](ship.md) only when authorized.
15
15
 
16
+ ## When a write outcome is uncertain
17
+
18
+ Use the existing storage and worker mechanisms; do not introduce a new platform for these rules.
19
+
20
+ - Before a replayable write, persist its tenant-scoped operation identity and payload identity, with enough state to recover after restart. Establish who owns an in-flight attempt so concurrent workers cannot independently replay it. Establish whether upstream deduplication is guaranteed, including its key, payload rules and retention window; sending a key alone proves nothing.
21
+ - A timeout, cancellation or lost response after dispatch may leave a completed side effect. Preserve that uncertainty across restart; stopping the caller is not rollback. Do not silently turn an uncertain attempt into a fresh operation.
22
+ - Reconcile against an authoritative receipt or lookup that matches the operation and payload. One verified result can confirm completion; conflicting or multiple matches require resolution. An empty stale, partial or eventually consistent lookup does not prove absence or authorize replay. Retry only under the verified deduplication contract or evidence establishing that repeating the write is safe.
23
+ - Keep unresolved attempts visible with safe error context, a next action and a known resolution owner, or an explicit ownership gap. Manual resolution still needs authority for any corrective write; do not manufacture completion to clear a queue.
24
+ - Test the relevant failure boundary: committed write with lost response, cancellation or restart before recording success, and stale lookup or concurrent replay where applicable. Record which were exercised and which remain unproven.
25
+
16
26
  ## Deliverable and acceptance
17
27
 
18
28
  Return the boundary contract, changed paths, environment, evidence at each tested layer, and remaining dependencies with owners when known. Done requires the agreed end-to-end result or an explicit narrower agreed scope. Do not silently replace live acceptance with a stub. In engagement mode, update the terrain/delivery record with confirmed facts; otherwise return the receipt directly.
@@ -3,10 +3,10 @@
3
3
  "version": 1,
4
4
  "files": {
5
5
  "SKILL.md": "7862670d5c023f0195e0b3fa8b52a3b33796813e53aaad083ae82a9a469e0faa",
6
- "references/build.md": "3dfeed619eeb1c8401f5cdf65e6f803fb209c70cb464dac4e60a1c890fd3a6f7",
6
+ "references/build.md": "6fa07b15b682119de53c38950f94942a58e4680ab58b24319d77c54a27cd9697",
7
7
  "references/debug.md": "c3bb344d38cc3552cb4e230c601a9be3fe173af2b2efb89aab6b7b04339f24f4",
8
8
  "references/eval-pack.md": "0590b85d3cae0903c6b1274540c92eaa2a4373047e8a0548d6942516ef0bb9e1",
9
- "references/integrate.md": "107a50bddf6cb0ba7f2bc006dfe9851800c6868e43f785a33ae4b737aeb74c95",
9
+ "references/integrate.md": "d0de35a783902ca8b4762e3a42a14f467766d56a928c3a5cf11adac2a6ba90ba",
10
10
  "references/qa.md": "d8f58e6d36436469a58aeb1107037f3e27fa81ff5b82d0e4df3c23eeadaf683c",
11
11
  "references/review.md": "63a007f78288089cc84cccc72647e8ce6721b7efa0f4f8d6774c0f0af594749d",
12
12
  "references/ship.md": "8cdcb2d4d6eb57e0adf3f1996bc02ae66920852ca304d2afd778fa483b7e969a",
@@ -6,7 +6,7 @@ Use the permitted context and authority in [task context](task-context.md). This
6
6
 
7
7
  ## Method
8
8
 
9
- 1. Identify the repository, its instructions, working tree, relevant callers, and test commands. Inspect examples before creating abstractions. Preserve unrelated edits and state which dependencies or interfaces the change touches.
9
+ 1. Identify the repository, its instructions, working tree, relevant callers, and test commands. Inspect examples before creating abstractions. Preserve unrelated edits and state which dependencies or interfaces the change touches. Before changing an untested legacy path, capture the undocumented behavior callers depend on with targeted characterization checks; distinguish behavior to preserve from the intended change.
10
10
  2. State the observable outcome, constraints, and acceptance checks. Reuse agreed criteria for routine fixes. If a consequential product choice is unresolved, surface that choice while continuing independent investigation; do not invent acceptance.
11
11
  3. Choose the smallest coherent path that demonstrates the outcome through the real entry point. Include the necessary storage, error handling, and interface behavior in that slice. Name the failure that stops expansion and the recovery path for stateful changes.
12
12
  4. Implement using the repository's tools and conventions. Search for existing services, fixtures, and validation before adding alternatives. Keep cleanup limited to what makes the changed path understandable; do not expand scope to repair unrelated code.
@@ -6,13 +6,23 @@ Start from [task context](task-context.md). Permitted supplied context is enough
6
6
 
7
7
  ## Method
8
8
 
9
- 1. Map producer, consumer, owner, direction, and side effects. Inspect the actual installed version and local implementation; verify uncertain behavior against current official documentation. Identify the relevant schema, authentication scopes, network boundary, and permitted test environment.
9
+ 1. Map producer, consumer, owner, direction, and side effects. Inspect the actual installed version and local implementation; verify uncertain behavior against current official documentation. Identify the relevant schema, authentication scopes, network boundary, and permitted test environment. Before changing an untested existing boundary, characterize the mappings, ordering or other observable behavior its callers depend on.
10
10
  2. Write the acceptance example: an input at the real boundary and the observable downstream result. Include a rejection or failure example. Separate configuration validity, successful authentication, transport connectivity, contract compatibility, and end-to-end behavior; none proves the next.
11
11
  3. Inspect credentials by presence and required scope without printing values. Use existing secret storage. Check data classification and retention before moving data; never pass raw `<private>` blocks into a model. Prefer sanitized or synthetic cases approved for the target environment.
12
- 4. Implement the narrow adapter using native repository patterns. Validate external inputs and model outputs, bound timeouts and retries, preserve error context without leaking payloads, and handle cancellation. For writes, establish idempotency or duplicate detection before retries; for events, check ordering, replay, and poison messages as applicable.
12
+ 4. Implement the narrow adapter using native repository patterns. Validate external inputs and model outputs, bound timeouts and retries, preserve error codes and failure phase without leaking payloads, and handle cancellation. Keep explicit authentication or permission rejections distinguishable from transport uncertainty. For writes, establish idempotency or duplicate detection before retries and apply the uncertain-write rules below when outcomes can be ambiguous; for events, check ordering, replay, and poison messages as applicable.
13
13
  5. Exercise a permitted success case and relevant failures: denied access, malformed data, rate limit, timeout, duplicate delivery, or partial completion. Trace correlation IDs or safe evidence across both sides. A mock proves client behavior only; if live access is unavailable, report that gap instead of claiming an integration works.
14
14
  6. Check cleanup and recovery for test side effects. Use [verification](verification.md) for receipts and [review](review.md) for security or data-contract changes. Route deployment through [ship](ship.md) only when authorized.
15
15
 
16
+ ## When a write outcome is uncertain
17
+
18
+ Use the existing storage and worker mechanisms; do not introduce a new platform for these rules.
19
+
20
+ - Before a replayable write, persist its tenant-scoped operation identity and payload identity, with enough state to recover after restart. Establish who owns an in-flight attempt so concurrent workers cannot independently replay it. Establish whether upstream deduplication is guaranteed, including its key, payload rules and retention window; sending a key alone proves nothing.
21
+ - A timeout, cancellation or lost response after dispatch may leave a completed side effect. Preserve that uncertainty across restart; stopping the caller is not rollback. Do not silently turn an uncertain attempt into a fresh operation.
22
+ - Reconcile against an authoritative receipt or lookup that matches the operation and payload. One verified result can confirm completion; conflicting or multiple matches require resolution. An empty stale, partial or eventually consistent lookup does not prove absence or authorize replay. Retry only under the verified deduplication contract or evidence establishing that repeating the write is safe.
23
+ - Keep unresolved attempts visible with safe error context, a next action and a known resolution owner, or an explicit ownership gap. Manual resolution still needs authority for any corrective write; do not manufacture completion to clear a queue.
24
+ - Test the relevant failure boundary: committed write with lost response, cancellation or restart before recording success, and stale lookup or concurrent replay where applicable. Record which were exercised and which remain unproven.
25
+
16
26
  ## Deliverable and acceptance
17
27
 
18
28
  Return the boundary contract, changed paths, environment, evidence at each tested layer, and remaining dependencies with owners when known. Done requires the agreed end-to-end result or an explicit narrower agreed scope. Do not silently replace live acceptance with a stub. In engagement mode, update the terrain/delivery record with confirmed facts; otherwise return the receipt directly.
@@ -3,7 +3,7 @@
3
3
  "version": 1,
4
4
  "files": {
5
5
  "SKILL.md": "c4152a7ee98027f810c3ab11403dea75dee5aae26da24f84de34630c1fd529b6",
6
- "references/earn-trust.md": "a3dc4b44686066fbc94fafa4b2fac3c5120d7940256898cec31279f565b912bb",
6
+ "references/earn-trust.md": "89df253665d996e28bc7acf785346c50f3062115ffd70fb7ad3300c99eb36ead",
7
7
  "references/task-context.md": "73eea2d7f164fac3226599e5be26ae4e79dcf69e0d24428dd12d623861410490"
8
8
  }
9
9
  }
@@ -2,52 +2,39 @@
2
2
 
3
3
  **Enter when:** new engagement where you don't have full access yet, trust is thin, the customer said "let's start small," or you need to navigate "we don't trust AI-generated code."
4
4
 
5
- **Read first:** `trust-profile.md`, `stakeholders.md`, `context.md`. The trust profile tells you where the walls are; the stakeholder map tells you who built them.
5
+ **Read first:** apply [task context](task-context.md), then retrieve permitted trust constraints, stakeholder evidence and current context. Never read raw private blocks.
6
6
 
7
- Trust is the currency of FDE work. Code quality gets you a second week; trust gets you the engagement. It's earned in small, visible moves - never demanded, never assumed, and never recovered once burned.
7
+ Build confidence through useful work, clear evidence and respect for the customer's process. Relationship confidence and access permissions are separate: a strong relationship does not grant production authority.
8
8
 
9
9
  ## Method (you do this work)
10
10
 
11
- **1. The trust ladder - every engagement climbs it in order:**
11
+ **1. Establish the access needed now.** Identify the next task, the minimum relevant access, and its actual policy or authorization source. Read-only access, reviewed PRs, branch writes and deployment rights can be granted independently. Reuse permissions already granted for the same scope; do not impose a ladder or ask the customer to re-earn established access. Record missing or disputed rights and continue work that does not depend on them.
12
12
 
13
- ```
14
- Level 0: Observer → read-only access, watching
15
- Level 1: Advisor → recommendations, no code changes
16
- Level 2: Contributor → PRs reviewed by their team
17
- Level 3: Committer → direct push to feature branches
18
- Level 4: Owner → production access, deploy authority
19
- Level 5: Trusted → they call you before making decisions
20
- ```
21
-
22
- **Never skip a level.** The FDE who asks for production access on day two gets observer access for a month. The FDE who ships a clean PR on day two gets committer access by week two. Each level is earned by demonstrating competence AND respect at the current level.
13
+ **2. Make progress visible.** Choose useful actions for the engagement's stage and agreed cadence:
23
14
 
24
- **2. The first-week trust plays - specific, not generic:**
25
-
26
- | Day | Move | Why it works |
27
- |-----|------|-------------|
28
- | 1 | Fix a small, visible, annoying bug - something the team has been stepping over (only after `success.md` has a signer, or the FDE overrides with the unknown still visible) | Proves you can ship in their environment without breaking things |
29
- | 1 | Ask the passed-over team what naming conventions they use - then use them | Shows respect before competence |
30
- | 2 | Send a one-paragraph status to the sponsor without being asked | Sets the pattern: they hear from you before they have to ask |
31
- | 3 | Find a genuine risk and flag it without drama | Demonstrates you're protecting them, not performing |
32
- | 5 | Show a small win to the champion so they can share it upward | Gives them evidence their bet on you was right |
15
+ - Deliver a small verified result within scope, or clarify a consequential unknown when implementation is premature. A quick win does not bypass release gates.
16
+ - Ask the existing team about conventions and prior attempts; credit their contributions without assuming they were passed over.
17
+ - Prepare a concise status update using the agreed channel and audience. Send only within existing communication authority.
18
+ - Flag a supported risk and its consequence without exaggerating urgency.
19
+ - Show results the intended users can evaluate, distinguishing demonstrated behavior from reported satisfaction.
33
20
 
34
21
  **3. Navigate "we don't trust AI-generated code":**
35
22
 
36
23
  This is increasingly common. The right response is respect, not persuasion:
37
24
 
38
25
  - **Ask the policy, don't assume.** "Does your organisation have a position on AI-assisted code in production?"
39
- - **If prohibited:** work without AI on their code. Use fdeops for engagement memory (`.fde/` files) and your own planning - that's your tooling, not theirs.
26
+ - **If prohibited:** do not load or work on their code with the model. Engagement notes and planning may also contain restricted data; use FDEOps on them only when that use is permitted. Continue with generic or explicitly permitted material, and identify what must be handled outside the AI workflow.
40
27
  - **If permitted with review:** every AI-touched line goes through their normal review process. Flag it: "AI-assisted, human-reviewed" in commit messages if they want traceability.
41
- - **If grey area:** treat as prohibited until someone with authority says otherwise. The cost of asking is zero; the cost of guessing wrong is the engagement.
42
- - **Never hide it.** An FDE caught using prohibited AI tools loses the engagement and the reputation. Full stop.
28
+ - **If grey area:** treat as prohibited until someone with authority says otherwise. Clarify only the policy needed for the next action and continue work on already permitted material.
29
+ - **Never hide it.** Disclose AI involvement according to the agreed policy; do not represent prohibited use as ordinary local tooling.
43
30
 
44
31
  **4. Trust recovery - when you've made a mistake:**
45
32
 
46
33
  Mistakes happen. What matters is speed and honesty:
47
34
 
48
- - **Own it in the first hour.** Not "we found an issue" - "I introduced this bug." Passive voice erodes trust faster than the mistake.
35
+ - **Report promptly under the incident process.** State the known impact and your confirmed contribution. Do not assign yourself or another person a cause before evidence supports it.
49
36
  - **Show the fix AND the prevention.** "Here's what happened, here's the fix, here's the test that prevents it next time."
50
- - **One visible win within 48 hours.** Trust recovery needs a concrete success close to the mistake - not weeks later.
37
+ - **Agree a recovery checkpoint.** Use a verified corrective result and a realistic next update; do not promise a win within an arbitrary window.
51
38
  - **Never minimise.** "It was a small bug" is your assessment, not theirs. Let them size it.
52
39
 
53
40
  **5. The trust account - deposits and withdrawals:**
@@ -65,36 +52,30 @@ Mistakes happen. What matters is speed and honesty:
65
52
 
66
53
  **`trust-profile.md`** - updated sections:
67
54
  ```markdown
68
- ## Trust level
69
- Current: <level 0-5> as of <date>
70
- Evidence: <what earned this level>
71
- Next target: <level> - requires: <specific action>
55
+ ## Access and working agreement
56
+ Current: <permitted task, system/environment and limits>
57
+ Source: <actual authorization/policy and date>
58
+ Next need: <access gap or none; responsible decision-maker if known>
72
59
 
73
60
  ## AI policy
74
61
  Status: <prohibited / permitted-with-review / grey-area-treating-as-prohibited>
75
62
  Source: <who confirmed, when>
76
63
  ```
77
64
 
78
- **`decisions.md`** - log trust-significant moves: "Flagged migration risk to ops lead before they discovered it (Day 3) - trust deposit."
65
+ **`decisions.md`** - record consequential confirmed agreements with their sources. A risk raised is an observed action; increased trust is not established unless supported by the customer's response.
79
66
 
80
67
  ## Checkpoint
81
68
 
82
- One question to the FDE: "Are we at the right trust level for what we need to do next week?" If not: name the gap, name the move, and put it in `context.md` as the next action.
83
-
84
- ## The week 2-4 valley
69
+ Check whether the next task has the required access and agreement. Reuse current evidence; if a material gap remains, name the applicable decision and next action.
85
70
 
86
- Week 1 is the honeymoon - everyone's excited, access is fresh, the brief is new. Weeks 2-4 are the valley: novelty wears off, real problems surface, the sponsor's patience shifts from "take your time" to "when do we see results." Most engagements silently fail here, not at ship.
71
+ ## Keep expectations current
87
72
 
88
- Counter it:
89
- - Ship one visible artifact per week, even if discovery isn't done. A terrain map, a risk register, a stakeholder signal update - something the sponsor can point to.
90
- - Proactive status update at end of week 2 - explicitly name what discovery revealed that wasn't in the brief. This resets expectations with evidence.
91
- - If still in discovery at week 3: the conversation with the sponsor about scope or timeline reset is overdue. Don't wait for them to ask.
73
+ At the agreed checkpoints, explain what changed, what has been demonstrated and what remains uncertain. If discovery invalidates the expected scope or timeline, surface the evidence when it affects the next decision. A long discovery phase may be appropriate for the work; week numbers alone do not establish impatience or failure.
92
74
 
93
75
  ## Principles
94
76
 
95
- - Trust is earned in small moves, lost in one. Never skip the ladder.
96
- - The first-week plays are specific and deliberate - not "be helpful."
97
- - AI policy: ask, never assume. Prohibited until confirmed.
98
- - Mistakes happen; hiding them doesn't. Own it in the first hour.
99
- - The FDE who makes the internal team look right earns trust faster than the FDE who ships the most code.
100
- - Weeks 2-4 are where engagements silently die. Ship visible artifacts weekly to survive the valley.
77
+ - Earn confidence through observable work; permissions come from applicable authority.
78
+ - Use the customer's conventions and credit actual contributions.
79
+ - AI policy applies to the data and use, including engagement memory.
80
+ - Report mistakes promptly, distinguish known causes from hypotheses, and verify recovery.
81
+ - Choose updates and follow-up timing from impact and the agreed cadence.
@@ -3,10 +3,10 @@
3
3
  "version": 1,
4
4
  "files": {
5
5
  "SKILL.md": "6a8f9490294e1a055c25984955d011db23ebb1eb2fb4f8634584d1455331425e",
6
- "references/build.md": "3dfeed619eeb1c8401f5cdf65e6f803fb209c70cb464dac4e60a1c890fd3a6f7",
6
+ "references/build.md": "6fa07b15b682119de53c38950f94942a58e4680ab58b24319d77c54a27cd9697",
7
7
  "references/debug.md": "c3bb344d38cc3552cb4e230c601a9be3fe173af2b2efb89aab6b7b04339f24f4",
8
8
  "references/eval-pack.md": "0590b85d3cae0903c6b1274540c92eaa2a4373047e8a0548d6942516ef0bb9e1",
9
- "references/integrate.md": "107a50bddf6cb0ba7f2bc006dfe9851800c6868e43f785a33ae4b737aeb74c95",
9
+ "references/integrate.md": "d0de35a783902ca8b4762e3a42a14f467766d56a928c3a5cf11adac2a6ba90ba",
10
10
  "references/qa.md": "d8f58e6d36436469a58aeb1107037f3e27fa81ff5b82d0e4df3c23eeadaf683c",
11
11
  "references/review.md": "63a007f78288089cc84cccc72647e8ce6721b7efa0f4f8d6774c0f0af594749d",
12
12
  "references/ship.md": "8cdcb2d4d6eb57e0adf3f1996bc02ae66920852ca304d2afd778fa483b7e969a",
@@ -6,7 +6,7 @@ Use the permitted context and authority in [task context](task-context.md). This
6
6
 
7
7
  ## Method
8
8
 
9
- 1. Identify the repository, its instructions, working tree, relevant callers, and test commands. Inspect examples before creating abstractions. Preserve unrelated edits and state which dependencies or interfaces the change touches.
9
+ 1. Identify the repository, its instructions, working tree, relevant callers, and test commands. Inspect examples before creating abstractions. Preserve unrelated edits and state which dependencies or interfaces the change touches. Before changing an untested legacy path, capture the undocumented behavior callers depend on with targeted characterization checks; distinguish behavior to preserve from the intended change.
10
10
  2. State the observable outcome, constraints, and acceptance checks. Reuse agreed criteria for routine fixes. If a consequential product choice is unresolved, surface that choice while continuing independent investigation; do not invent acceptance.
11
11
  3. Choose the smallest coherent path that demonstrates the outcome through the real entry point. Include the necessary storage, error handling, and interface behavior in that slice. Name the failure that stops expansion and the recovery path for stateful changes.
12
12
  4. Implement using the repository's tools and conventions. Search for existing services, fixtures, and validation before adding alternatives. Keep cleanup limited to what makes the changed path understandable; do not expand scope to repair unrelated code.
@@ -6,13 +6,23 @@ Start from [task context](task-context.md). Permitted supplied context is enough
6
6
 
7
7
  ## Method
8
8
 
9
- 1. Map producer, consumer, owner, direction, and side effects. Inspect the actual installed version and local implementation; verify uncertain behavior against current official documentation. Identify the relevant schema, authentication scopes, network boundary, and permitted test environment.
9
+ 1. Map producer, consumer, owner, direction, and side effects. Inspect the actual installed version and local implementation; verify uncertain behavior against current official documentation. Identify the relevant schema, authentication scopes, network boundary, and permitted test environment. Before changing an untested existing boundary, characterize the mappings, ordering or other observable behavior its callers depend on.
10
10
  2. Write the acceptance example: an input at the real boundary and the observable downstream result. Include a rejection or failure example. Separate configuration validity, successful authentication, transport connectivity, contract compatibility, and end-to-end behavior; none proves the next.
11
11
  3. Inspect credentials by presence and required scope without printing values. Use existing secret storage. Check data classification and retention before moving data; never pass raw `<private>` blocks into a model. Prefer sanitized or synthetic cases approved for the target environment.
12
- 4. Implement the narrow adapter using native repository patterns. Validate external inputs and model outputs, bound timeouts and retries, preserve error context without leaking payloads, and handle cancellation. For writes, establish idempotency or duplicate detection before retries; for events, check ordering, replay, and poison messages as applicable.
12
+ 4. Implement the narrow adapter using native repository patterns. Validate external inputs and model outputs, bound timeouts and retries, preserve error codes and failure phase without leaking payloads, and handle cancellation. Keep explicit authentication or permission rejections distinguishable from transport uncertainty. For writes, establish idempotency or duplicate detection before retries and apply the uncertain-write rules below when outcomes can be ambiguous; for events, check ordering, replay, and poison messages as applicable.
13
13
  5. Exercise a permitted success case and relevant failures: denied access, malformed data, rate limit, timeout, duplicate delivery, or partial completion. Trace correlation IDs or safe evidence across both sides. A mock proves client behavior only; if live access is unavailable, report that gap instead of claiming an integration works.
14
14
  6. Check cleanup and recovery for test side effects. Use [verification](verification.md) for receipts and [review](review.md) for security or data-contract changes. Route deployment through [ship](ship.md) only when authorized.
15
15
 
16
+ ## When a write outcome is uncertain
17
+
18
+ Use the existing storage and worker mechanisms; do not introduce a new platform for these rules.
19
+
20
+ - Before a replayable write, persist its tenant-scoped operation identity and payload identity, with enough state to recover after restart. Establish who owns an in-flight attempt so concurrent workers cannot independently replay it. Establish whether upstream deduplication is guaranteed, including its key, payload rules and retention window; sending a key alone proves nothing.
21
+ - A timeout, cancellation or lost response after dispatch may leave a completed side effect. Preserve that uncertainty across restart; stopping the caller is not rollback. Do not silently turn an uncertain attempt into a fresh operation.
22
+ - Reconcile against an authoritative receipt or lookup that matches the operation and payload. One verified result can confirm completion; conflicting or multiple matches require resolution. An empty stale, partial or eventually consistent lookup does not prove absence or authorize replay. Retry only under the verified deduplication contract or evidence establishing that repeating the write is safe.
23
+ - Keep unresolved attempts visible with safe error context, a next action and a known resolution owner, or an explicit ownership gap. Manual resolution still needs authority for any corrective write; do not manufacture completion to clear a queue.
24
+ - Test the relevant failure boundary: committed write with lost response, cancellation or restart before recording success, and stale lookup or concurrent replay where applicable. Record which were exercised and which remain unproven.
25
+
16
26
  ## Deliverable and acceptance
17
27
 
18
28
  Return the boundary contract, changed paths, environment, evidence at each tested layer, and remaining dependencies with owners when known. Done requires the agreed end-to-end result or an explicit narrower agreed scope. Do not silently replace live acceptance with a stub. In engagement mode, update the terrain/delivery record with confirmed facts; otherwise return the receipt directly.