@mmerterden/multi-agent-pipeline 14.1.1 → 14.2.1

This diff represents the content of publicly available package versions that have been released to one of the supported registries. The information contained in this diff is provided for informational purposes only and reflects changes between package versions as they appear in their respective public registries.
Files changed (34) hide show
  1. package/CHANGELOG.md +120 -0
  2. package/package.json +1 -1
  3. package/pipeline/commands/deploy.md +4 -1
  4. package/pipeline/commands/multi-agent/SKILL.md +6 -3
  5. package/pipeline/commands/multi-agent/help/SKILL.md +49 -11
  6. package/pipeline/commands/multi-agent/setup/SKILL.md +1 -1
  7. package/pipeline/commands/multi-agent/store-ready/SKILL.md +340 -0
  8. package/pipeline/commands/multi-agent/sync/SKILL.md +4 -3
  9. package/pipeline/commands/multi-agent/test/SKILL.md +18 -8
  10. package/pipeline/commands/multi-agent/test-accessibility/SKILL.md +33 -0
  11. package/pipeline/commands/multi-agent/test-dark-mode/SKILL.md +33 -0
  12. package/pipeline/commands/multi-agent/test-dynamic-type/SKILL.md +33 -0
  13. package/pipeline/commands/multi-agent/test-screenshots/SKILL.md +41 -0
  14. package/pipeline/commands/multi-agent/testflight-validation/SKILL.md +28 -201
  15. package/pipeline/commands/sim-test.md +45 -36
  16. package/pipeline/multi-agent-refs/cross-cli-contract.md +3 -2
  17. package/pipeline/multi-agent-refs/knowledge.md +1 -1
  18. package/pipeline/multi-agent-refs/phases/phase-0-init.md +1 -1
  19. package/pipeline/schemas/prefs.schema.json +1 -1
  20. package/pipeline/skills/.skills-index.json +48 -3
  21. package/pipeline/skills/shared/README.md +11 -6
  22. package/pipeline/skills/shared/core/multi-agent/SKILL.md +13 -17
  23. package/pipeline/skills/shared/core/multi-agent-help/SKILL.md +50 -12
  24. package/pipeline/skills/shared/core/multi-agent-purge/SKILL.md +18 -3
  25. package/pipeline/skills/shared/core/multi-agent-store-ready/SKILL.md +50 -0
  26. package/pipeline/skills/shared/core/multi-agent-sync/SKILL.md +4 -3
  27. package/pipeline/skills/shared/core/multi-agent-test/SKILL.md +18 -8
  28. package/pipeline/skills/shared/core/multi-agent-test-accessibility/SKILL.md +37 -0
  29. package/pipeline/skills/shared/core/multi-agent-test-dark-mode/SKILL.md +37 -0
  30. package/pipeline/skills/shared/core/multi-agent-test-dynamic-type/SKILL.md +37 -0
  31. package/pipeline/skills/shared/core/multi-agent-test-screenshots/SKILL.md +44 -0
  32. package/pipeline/skills/shared/core/multi-agent-testflight-validation/SKILL.md +29 -101
  33. package/pipeline/skills/shared/external/firebase/SKILL.md +1 -1
  34. package/pipeline/skills/skills-index.md +8 -3
@@ -59,7 +59,7 @@ Run every step automatically:
59
59
  ```
60
60
  Step 1: PLATFORM Detect macOS / Linux / Windows (Git Bash / WSL); export PLATFORM env
61
61
  Step 1.5: DETECT Compare timestamps, find stale targets
62
- Step 2: COPILOT Claude Code -> Copilot CLI (instructions + 44 sub-command skills)
62
+ Step 2: COPILOT Claude Code -> Copilot CLI (instructions + 49 sub-command skills)
63
63
  Step 2b: CODEX Claude Code -> Codex CLI (1 router skill + 44 specs as refs + 8 agent TOML)
64
64
  Step 3: REPO Claude Code -> pipeline repo (genericized, personal data scrub, bash -n on all sh)
65
65
  Step 3c: PLUGINS pipeline shared/external -> multi-agent-plugins marketplace (rebuild knowledge/,
@@ -491,14 +491,15 @@ same 43 specs as reference files rather than as peer skills, via Step 2b - see
491
491
  |-------------|-------------|
492
492
  | `~/.claude/commands/multi-agent/{cmd}/SKILL.md` | `~/.copilot/skills/multi-agent-{cmd}/SKILL.md` |
493
493
 
494
- **44 commands are synced** (canonical inventory - must match `cross-cli-contract.md` section 1; drift = contract violation):
494
+ **49 commands are synced** (canonical inventory - must match `cross-cli-contract.md` section 1; drift = contract violation):
495
495
 
496
496
  ```
497
497
  analysis, analysis-resolve, autopilot, build-optimize, channels, create-jira, design-check, dev,
498
498
  dev-autopilot, dev-local, dev-local-autopilot, diff-explain, forget, garbage-collect,
499
499
  help, ios-coding-standard, issue, jira, kill, language, local,
500
500
  local-autopilot, log, manual-test, prune-logs, purge, refactor, resume, review, review-issue, review-jira,
501
- routines, save, scan, search, setup, ship, stack, status, sync, test, testflight-validation, uninstall, update
501
+ routines, save, scan, search, setup, ship, stack, status, store-ready, sync, test, test-accessibility,
502
+ test-dark-mode, test-dynamic-type, test-screenshots, testflight-validation, uninstall, update
502
503
  ```
503
504
 
504
505
  **NOT synced**: `$HOME/.claude/multi-agent-refs/*` - lazy-load references, Claude Code specific
@@ -22,14 +22,24 @@ Pass `$ARGUMENTS` through verbatim - sim-test.md decides which test matrix to
22
22
 
23
23
  ## Quick reference
24
24
 
25
- | Invocation | What it does |
26
- |---|---|
27
- | `/multi-agent:test` | Walk every screen, collect screenshots + UI tree, produce a general report |
28
- | `/multi-agent:test "dark mode"` | Light vs dark comparison, contrast + colour bugs |
29
- | `/multi-agent:test "accessibility"` | VoiceOver label + tap-target + contrast audit |
30
- | `/multi-agent:test "dynamic type"` | Layout check at XL-XXXL text sizes |
31
- | `/multi-agent:test "screenshot tr"` | App Store screenshot set (Turkish locale) |
32
- | `/multi-agent:test "store-ready"` | App Store guideline pre-flight check |
25
+ | Invocation | Fixed-scenario alias | What it does |
26
+ |---|---|---|
27
+ | `/multi-agent:test` | - | Walk every screen, collect screenshots + UI tree, produce a general report |
28
+ | `/multi-agent:test "dark mode"` | `/multi-agent:test-dark-mode` | Light vs dark comparison, contrast + colour bugs |
29
+ | `/multi-agent:test "accessibility"` | `/multi-agent:test-accessibility` | VoiceOver label + tap-target + contrast audit |
30
+ | `/multi-agent:test "dynamic type"` | `/multi-agent:test-dynamic-type` | Layout check at XL-XXXL text sizes |
31
+ | `/multi-agent:test "screenshot <lang>"` | `/multi-agent:test-screenshots [locale]` | App Store screenshot set in a locale (alias defaults to tr) |
32
+ | `/multi-agent:test "store-ready" [path]` | hands off to `/multi-agent:store-ready` | package validation, not a UI test |
33
+
34
+ Both columns are supported and neither is deprecated; the aliases exist so the
35
+ scenario list autocompletes off `test-` instead of having to be remembered.
36
+
37
+ `store-ready` is the odd one out: it validates a built **package** rather than a
38
+ running app, so it is not implemented here at all. `/multi-agent:store-ready` owns
39
+ it on both platforms - three gates per platform, plus this file's sweep as its
40
+ Step A. `/multi-agent:testflight-validation` is the iOS-pinned alias of that
41
+ command. The quoted tag is kept as a hand-off so an existing invocation still lands
42
+ somewhere correct.
33
43
 
34
44
  ## Requirements
35
45
 
@@ -0,0 +1,33 @@
1
+ ---
2
+ description: "Accessibility audit on a booted simulator / emulator: VoiceOver labels, sub-44pt tap targets, contrast, traits. Alias pinning the accessibility scenario. Use when auditing a screen for assistive-technology support."
3
+ description-tr: "Booted simulator / emulator üzerinde erişilebilirlik denetimi: VoiceOver label'ları, 44x44pt altı tap target'lar, kontrast oranları ve trait'ler; her ekranda canlı UI tree'ye karşı denetlenir. /multi-agent:test \"accessibility\" için sabit-senaryo alias'ı."
4
+ argument-hint: "(argümansız - senaryo sabit)"
5
+ ---
6
+
7
+ # /multi-agent:test-accessibility - Accessibility audit
8
+
9
+ Fixed-scenario alias. The implementation is the UI Bug Hunter flow; this command
10
+ only pins the scenario so the tag does not have to be typed or quoted.
11
+
12
+ ## Dispatcher
13
+
14
+ Read this file and apply its instructions:
15
+
16
+ ```
17
+ $HOME/.claude/commands/sim-test.md
18
+ ```
19
+
20
+ Run it with **scenario = `accessibility`**. Ignore `$ARGUMENTS`; the scenario is
21
+ fixed by the command name. Anything the user adds is context for the report, never
22
+ a scenario override - to run a different matrix they invoke that command.
23
+
24
+ ## Equivalent
25
+
26
+ `/multi-agent:test "accessibility"` - identical behaviour. Both forms are
27
+ supported; neither is deprecated.
28
+
29
+ ## Requirements
30
+
31
+ - **iOS**: Xcode + a booted Simulator (`xcrun simctl list | grep Booted`)
32
+ - **Android**: Android SDK + a running emulator or USB device (`adb devices`)
33
+ - **MCP**: the `dev-toolkit` MCP server must be installed (tool names start with `mcp__dev-toolkit__`)
@@ -0,0 +1,33 @@
1
+ ---
2
+ description: "Dark mode UI test on a booted simulator / emulator: walk every screen light then dark, report contrast + colour bugs. Alias pinning the dark-mode scenario. Use when a dark-mode rendering bug is suspected."
3
+ description-tr: "Booted simulator / emulator üzerinde dark mode UI testi. Her ekranı light'ta gezer, görünümü değiştirir, tekrar gezer; farktan kontrast + renk + sabit-palet buglarını raporlar. /multi-agent:test \"dark mode\" için sabit-senaryo alias'ı."
4
+ argument-hint: "(argümansız - senaryo sabit)"
5
+ ---
6
+
7
+ # /multi-agent:test-dark-mode - Dark mode UI test
8
+
9
+ Fixed-scenario alias. The implementation is the UI Bug Hunter flow; this command
10
+ only pins the scenario so the tag does not have to be typed or quoted.
11
+
12
+ ## Dispatcher
13
+
14
+ Read this file and apply its instructions:
15
+
16
+ ```
17
+ $HOME/.claude/commands/sim-test.md
18
+ ```
19
+
20
+ Run it with **scenario = `dark mode`**. Ignore `$ARGUMENTS`; the scenario is fixed
21
+ by the command name. Anything the user adds is context for the report, never a
22
+ scenario override - to run a different matrix they invoke that command.
23
+
24
+ ## Equivalent
25
+
26
+ `/multi-agent:test "dark mode"` - identical behaviour. Both forms are supported;
27
+ neither is deprecated.
28
+
29
+ ## Requirements
30
+
31
+ - **iOS**: Xcode + a booted Simulator (`xcrun simctl list | grep Booted`)
32
+ - **Android**: Android SDK + a running emulator or USB device (`adb devices`)
33
+ - **MCP**: the `dev-toolkit` MCP server must be installed (tool names start with `mcp__dev-toolkit__`)
@@ -0,0 +1,33 @@
1
+ ---
2
+ description: "Dynamic Type layout test on a booted simulator / emulator: re-walk every screen at XL through accessibility-extra-large, report truncation and clipping. Alias pinning the dynamic-type scenario. Use when checking a layout survives large text."
3
+ description-tr: "Booted simulator / emulator üzerinde Dynamic Type layout testi. Her ekranı XL'den accessibility-extra-large'a kadar yeniden gezer; kırpılma, taşma, üst üste binme ve ölçeklenmeyi bırakan sabit yükseklikli container'ları raporlar. /multi-agent:test \"dynamic type\" için sabit-senaryo alias'ı."
4
+ argument-hint: "(argümansız - senaryo sabit)"
5
+ ---
6
+
7
+ # /multi-agent:test-dynamic-type - Dynamic Type layout test
8
+
9
+ Fixed-scenario alias. The implementation is the UI Bug Hunter flow; this command
10
+ only pins the scenario so the tag does not have to be typed or quoted.
11
+
12
+ ## Dispatcher
13
+
14
+ Read this file and apply its instructions:
15
+
16
+ ```
17
+ $HOME/.claude/commands/sim-test.md
18
+ ```
19
+
20
+ Run it with **scenario = `dynamic type`**. Ignore `$ARGUMENTS`; the scenario is
21
+ fixed by the command name. Anything the user adds is context for the report, never
22
+ a scenario override - to run a different matrix they invoke that command.
23
+
24
+ ## Equivalent
25
+
26
+ `/multi-agent:test "dynamic type"` - identical behaviour. Both forms are
27
+ supported; neither is deprecated.
28
+
29
+ ## Requirements
30
+
31
+ - **iOS**: Xcode + a booted Simulator (`xcrun simctl list | grep Booted`)
32
+ - **Android**: Android SDK + a running emulator or USB device (`adb devices`)
33
+ - **MCP**: the `dev-toolkit` MCP server must be installed (tool names start with `mcp__dev-toolkit__`)
@@ -0,0 +1,41 @@
1
+ ---
2
+ description: "App Store screenshot set from a booted simulator / emulator in a given locale, taken as an argument and defaulting to tr. Alias pinning the screenshot scenario. Use when generating store screenshots or checking a localised build renders."
3
+ description-tr: "Booted simulator / emulator'dan belirtilen dilde App Store screenshot seti: temiz status bar, her ekran gezilip yakalanır. Dili isim içine gömmek yerine argüman olarak alır (default tr). /multi-agent:test \"screenshot <lang>\" için sabit-senaryo alias'ı."
4
+ argument-hint: "[locale] - boş = tr; örn. tr | en | de | ar"
5
+ ---
6
+
7
+ # /multi-agent:test-screenshots - Locale screenshot set
8
+
9
+ Fixed-scenario alias, **parameterised by locale**. The scenario is pinned; the
10
+ language is not, because one command per language would put a parameter in a name.
11
+
12
+ ## Locale resolution
13
+
14
+ 1. `$ARGUMENTS` names a locale (`tr`, `en`, `de`, `ar`, ...) → use it.
15
+ 2. `$ARGUMENTS` empty → `tr`.
16
+ 3. `$ARGUMENTS` is something other than a locale → do not guess a language and do
17
+ not silently fall back. Ask which locale with `AskUserQuestion`, rendered in
18
+ `outputLanguage`.
19
+
20
+ ## Dispatcher
21
+
22
+ Read this file and apply its instructions:
23
+
24
+ ```
25
+ $HOME/.claude/commands/sim-test.md
26
+ ```
27
+
28
+ Run it with **scenario = `screenshot <locale>`**, substituting the locale resolved
29
+ above. That is the tag sim-test.md matches; passing a bare `screenshot` leaves the
30
+ language unset.
31
+
32
+ ## Equivalent
33
+
34
+ `/multi-agent:test "screenshot tr"` - identical behaviour. Both forms are
35
+ supported; neither is deprecated.
36
+
37
+ ## Requirements
38
+
39
+ - **iOS**: Xcode + a booted Simulator (`xcrun simctl list | grep Booted`)
40
+ - **Android**: Android SDK + a running emulator or USB device (`adb devices`)
41
+ - **MCP**: the `dev-toolkit` MCP server must be installed (tool names start with `mcp__dev-toolkit__`)
@@ -1,219 +1,46 @@
1
1
  ---
2
- description: "Pre-submission validation for a TestFlight / App Store build (iOS, local-only). Three gates: static archive audit, Apple's own `altool --validate-app`, and a Review-Guidelines check. ITMS codes are mapped to the rule each implies. Validates only, never uploads. Use when a build is about to go to TestFlight, or a submission was rejected and you need why."
3
- description-tr: "TestFlight / App Store yüklemesi öncesi doğrulama (iOS, yalnızca lokal). Repo + branch seç, sonra ya build'i sen ver ya da koşu archive alsın; üç kapıyı geç: statik 18-kurallı archive denetimi, Apple'ın kendi `altool --validate-app`'i, ve App Store Review Guidelines'a karşı guideline incelemesi. Her kapı için ayrı verdict, ITMS kodları ilgili kurala eşlenmiş. Sadece doğrular - asla yüklemez."
2
+ description: "iOS-pinned alias of store-ready: same three gates on an iOS build, one implementation. Validates a TestFlight / App Store package, never uploads. Use when an iOS build is about to go to TestFlight, or a submission was rejected and you need why."
3
+ description-tr: "store-ready komutunun iOS'a sabitlenmiş alias'ı: iOS build'inde aynı üç kapı, tek implementasyon. TestFlight / App Store paketini doğrular, asla yüklemez."
4
4
  argument-hint: "[repo] - empty = pick from prefs; repo name or path; --ipa=<path>; --archive=<path>; --resume"
5
5
  allowed-tools: Agent, Bash, Read, Write, Edit, Glob, Grep, TaskCreate, TaskUpdate, AskUserQuestion, Skill, mcp__dev-toolkit__ios_app_store_audit, mcp__dev-toolkit__ios_export_ipa, mcp__dev-toolkit__ios_testflight_validate, mcp__dev-toolkit__ios_xcodebuild, mcp__dev-toolkit__ios_xcresult
6
6
  ---
7
7
 
8
- # /multi-agent:testflight-validation - pre-submission validation
8
+ # /multi-agent:testflight-validation - iOS alias for store-ready
9
9
 
10
- Catch, before you upload, what App Store Connect would send back after you do.
10
+ This command is an alias. The implementation it used to carry was merged into
11
+ `/multi-agent:store-ready`, which runs the same three gates on iOS and adds the
12
+ Android side, so there is one flow to maintain instead of two that had already
13
+ started to drift.
11
14
 
12
- **Local-only.** No commits, no push, no PR, no channels. The worktree exists only
13
- to archive without touching your working tree.
15
+ The name is kept because it is the one people reach for when the target is
16
+ TestFlight, and removing a command is a breaking change to the slash-command
17
+ surface.
14
18
 
15
- **It never uploads.** Only `--validate-app` is ever invoked, never `--upload-app`.
16
- A validation run must not be able to ship a build by accident.
19
+ ## Dispatcher
17
20
 
18
- ## Why three gates and not one
19
-
20
- Each gate sees something the others structurally cannot. Reporting one of them as
21
- "the check" is how a build passes locally and gets rejected anyway.
22
-
23
- | Gate | What runs | Needs | Sees | Blind to |
24
- |---|---|---|---|---|
25
- | **1. Static** | `ios_app_store_audit` (18 rules, real ITMS codes) | an `.xcarchive` | privacy manifest, required-reason API, Info.plist, code signing, entitlements, embedded SDK, IPv6, debug-tool leak, binary size | anything that depends on the App Store Connect account |
26
- | **2. Authoritative** | `ios_testflight_validate` → `altool --validate-app` | an `.ipa` + credentials | unregistered bundle ID, profile that does not match the app record, **a version+build pair already used**, entitlements not provisioned for the App ID | the Review Guidelines - Apple's validator does not read them |
27
- | **3. Guideline** | `app-store-review` skill + repo evidence | repo checkout | ATT flow, privacy policy, account deletion, IAP rules, purpose-string wording, permission justification | anything not visible in source |
28
-
29
- Gate 2 is the only one that asks Apple, and Gate 3 is the only one that covers the
30
- rejections a human reviewer writes. Most "we passed validation and still got
31
- rejected" cases are Gate 3 findings.
32
-
33
- ## Step 0 - parse input
34
-
35
- | Input | Meaning |
36
- |---|---|
37
- | (empty) | ask which repo (Step 1) |
38
- | `my-ios-app` or a path | that repo |
39
- | `--ipa=<path>` | Mode B with an `.ipa`; Gate 1 cannot run (see Step 3) |
40
- | `--archive=<path>` | Mode B with an `.xcarchive`; all three gates run |
41
- | `--resume` | continue the last run from its state file |
42
-
43
- State lives at `$HOME/.claude/logs/multi-agent/<task_id>/agent-state.json` with
44
- `taskId = TFV-<repo>-<yyyymmddHHMM>`. Register phases with the tracker
45
- (`$HOME/.claude/multi-agent-refs/tracker-contract.md`) so `:resume` and `:status`
46
- work like any other run.
47
-
48
- ## Step 1 - pickers (native, always)
49
-
50
- Use `AskUserQuestion` for every step - never a numbered text menu. Questions and
51
- descriptions render in `prefs.global.outputLanguage`; `label` and `header` stay
52
- English, per `$HOME/.claude/multi-agent-refs/picker-contract.md`. Print the
53
- `Step <i>/<n>: <what this decides>` breadcrumb for each.
54
-
55
- 1. **Repo** - from `prefs.projects` where the stack is iOS. A single match
56
- auto-resolves (say so in the breadcrumb, do not silently skip the step).
57
- 2. **Branch** - the branch to validate. Resolution order:
58
- - `git fetch --prune` first, capturing **stderr**. **If the fetch fails, do not
59
- silently fall back to a cached ref**, and **classify before naming a cause** -
60
- the same rule as the `/multi-agent:dev` remote gate:
61
-
62
- | stderr contains | Cause | Remedy |
63
- |---|---|---|
64
- | `could not read Password`, `Authentication failed`, `403` | credential | store the PAT in the credential helper or switch the remote to SSH. **A VPN cannot fix this**, and the base ref being stale is unrelated to what broke - do not offer the cached-ref fallback. |
65
- | `Could not resolve host`, `Operation timed out`, `Connection refused` | network | retry / continue on the cached ref with an explicit warning / switch remote / abort, per `$HOME/.claude/multi-agent-refs/rules.md` |
66
- | `Repository not found`, `404` | wrong remote | show `git remote -v` and ask |
67
-
68
- Always print the observed stderr line next to the classification. Asserting
69
- `unreachable (VPN/DNS)` for a missing-credential error that returns in under a
70
- second sends the user to fix something that was never broken.
71
- - Offer the current branch, the default branch, and any `release/*` /
72
- `tkdevelop/*` heads.
73
- 3. **Mode** - how the build is obtained:
74
- - `Supply a build` (Mode B, default) - fastest, no signing needed in-run.
75
- - `Archive from this branch` (Mode A) - needs a distribution certificate and
76
- profile in the keychain, and takes as long as a release archive.
77
-
78
- ## Step 2 - pre-flight, before anything expensive
79
-
80
- Report every line; a missing prerequisite is a halt, not a warning.
81
-
82
- ```bash
83
- xcrun --find altool >/dev/null 2>&1 || echo "MISSING: altool (install Xcode)"
84
- xcodebuild -version | head -1
85
- ```
86
-
87
- Then resolve credentials, and **state which tier is active in the report**:
88
-
89
- | Tier | Source | Effect |
90
- |---|---|---|
91
- | 1 | ASC API key - key id + issuer id from the keychain via `prefs.global.keychainMapping`, `.p8` at `~/.appstoreconnect/private_keys/AuthKey_<keyId>.p8` | Gate 2 runs |
92
- | 2 | Apple ID + app-specific password, referenced as a keychain item | Gate 2 runs |
93
- | 3 | neither | **Gate 2 reports `SKIPPED`, and the run says so in the verdict line** |
94
-
95
- Credentials come from `/multi-agent:setup`; never prompt for a secret value in
96
- chat. If nothing is configured, tell the user which of the two tiers they can set
97
- up and that tier 2 needs no elevated App Store Connect role.
98
-
99
- Multi-provider accounts need `--provider-public-id`. When it is not in prefs, run
100
- `ios_testflight_validate({list_providers: true})` once and ask which provider.
101
-
102
- ## Step 3 - obtain the build
103
-
104
- ### Mode B - a build you supply
105
-
106
- - `.xcarchive` → Gate 1 runs on it. To reach Gate 2 the archive must be exported,
107
- so run `ios_export_ipa` (see Mode A step 3 for the signing inputs).
108
- - `.ipa` only → **Gate 1 is reported `SKIPPED (needs .xcarchive)`.** The static
109
- audit reads archive structure that an `.ipa` does not carry. Do not present a
110
- two-gate run as a full pass; say which gate did not run and why, and offer to
111
- re-run with the archive.
112
-
113
- ### Mode A - archive from the branch
114
-
115
- 1. Worktree at `{projectRoot}/{worktreeBasePath}/{taskId}` on the chosen branch.
116
- **Never under `$HOME`**, never a direct checkout of the main working tree.
117
- 2. Resolve the scheme and workspace/project from prefs; ask if ambiguous.
118
- 3. Archive:
119
- `ios_xcodebuild({workspace|project, scheme, action: "archive", configuration: "Release", destination: "generic/platform=iOS"})`
120
- Note the destination: the simulator default would produce an archive that
121
- cannot be exported for distribution.
122
- 4. Export:
123
- `ios_export_ipa({archive_path, output_dir, method: "app-store-connect", team_id, provisioning_profiles?, signing_style?})`
124
- Leave `allow_provisioning_updates` off unless the user asks: it lets xcodebuild
125
- create or modify profiles in the developer account, which a validation run has
126
- no business doing.
127
- 5. A failed export halts with the parsed errors. The usual causes are a missing
128
- distribution certificate, a profile that does not match the bundle ID, or
129
- `signing_style: "manual"` with no `provisioning_profiles` map.
130
-
131
- ## Step 4 - Gate 1, static audit
132
-
133
- `ios_app_store_audit({archive_path, rules: "all"})`.
134
-
135
- `error` findings are blocking; `warning` is advisory. Group the output by severity
136
- and keep each finding's ITMS code - Gate 2 may return the same code, and seeing
137
- it in both places tells the user it is real rather than a heuristic.
138
-
139
- ## Step 5 - Gate 2, Apple's own validation
140
-
141
- `ios_testflight_validate({ipa_path, platform: "ios", <credential args>})`.
142
-
143
- Render the verdict exactly as returned:
144
-
145
- - `PASS` - Apple accepted the binary for delivery.
146
- - `FAIL` - list each issue with its ITMS code, the mapped guideline, and the hint.
147
- - `SKIPPED` - print the reason. **Never render this as a pass.** The verdict line
148
- for the whole run must read `2 of 3 gates cleared, 1 skipped`, not `passed`.
149
-
150
- Gate 2 is the only gate that catches a build number already used - the most
151
- common wasted upload - so when it fails on that, say so plainly and name the next
152
- free build number.
153
-
154
- ## Step 6 - Gate 3, guideline review
155
-
156
- Load the `app-store-review` skill and review the repo against it. This is the gate
157
- that catches what a human reviewer rejects, so it reads source, not the binary:
158
-
159
- | Area | Evidence to gather |
160
- |---|---|
161
- | Purpose strings | every `NS*UsageDescription` in Info.plist - present, specific, user-facing, and matching what the code actually does with the data |
162
- | Privacy manifest | `PrivacyInfo.xcprivacy` exists, declares required-reason APIs, and matches the SDKs actually linked |
163
- | Tracking | if any tracking API or SDK is present, an ATT prompt exists and runs before collection |
164
- | Account deletion | if the app creates accounts, an in-app deletion path exists (guideline 5.1.1(v)) |
165
- | Privacy policy | reachable in-app and in the metadata |
166
- | IAP | anything unlocking features goes through StoreKit, with no external purchase path |
167
- | Sign in with Apple | present when a third-party social login is offered |
168
-
169
- For each: `pass` / `fail` / `not-applicable` with the evidence path that justifies
170
- it. `not-applicable` needs a reason - an unexamined area is not a pass.
171
-
172
- ## Step 7 - report
173
-
174
- Write to `~/TestFlightChecks/<repo>-<branch>-<timestamp>/report.md` and print a
175
- summary. Structure:
21
+ Read this file and apply its instructions:
176
22
 
177
23
  ```
178
- Verdict: <N> of 3 gates cleared[, <M> skipped]
179
- Build: <ipa or archive path> · <bundle id> <version> (<build>)
180
- Auth: tier <1|2|none> · <method>
181
-
182
- Gate 1 static audit PASS | FAIL (<n> blocking, <n> advisory) | SKIPPED (<reason>)
183
- Gate 2 Apple validation PASS | FAIL (<n> issues) | SKIPPED (<reason>)
184
- Gate 3 guideline review PASS | FAIL (<n> findings) | <n> not-applicable
185
-
186
- Blocking - fix before uploading
187
- [ITMS-90683] Info.plist: NSCameraUsageDescription missing
188
- guideline 5.1.1 Data Collection and Storage
189
- <hint>
190
- <file:line>
191
-
192
- Advisory
193
- ...
194
-
195
- Not run
196
- Gate 1: needs an .xcarchive; only an .ipa was supplied
24
+ $HOME/.claude/commands/multi-agent/store-ready/SKILL.md
197
25
  ```
198
26
 
199
- Rules for the report:
27
+ Run it with **platform pinned to `ios`** and skip its platform-detection step. Pass
28
+ `$ARGUMENTS` through verbatim: `--ipa=`, `--archive=`, a repo name or path, and
29
+ `--resume` all mean exactly what they mean there.
200
30
 
201
- - A skipped gate is never folded into the pass count. The verdict line states the
202
- skip.
203
- - Every blocking finding carries a file path or an ITMS code. A finding the user
204
- cannot act on is noise.
205
- - No AI or assistant attribution anywhere, per
206
- `$HOME/.claude/rules/git-conventions.md`.
207
- - Real newlines, no HTML entities, per
208
- `$HOME/.claude/multi-agent-refs/rules.md "External System Outputs"`.
31
+ `--aab=` / `--apk=` are Android inputs and are not valid here. If one is supplied,
32
+ do not silently switch platform - say the Android inputs belong to
33
+ `/multi-agent:store-ready`, and stop.
209
34
 
210
- ## Step 8 - offer the next action, do not take it
35
+ ## What you get
211
36
 
212
- Print, and stop:
37
+ Unchanged from before the merge, on the iOS path:
213
38
 
214
- - the exact `xcrun altool --upload-app` command for when the gates are clear, so
215
- uploading stays an explicit human act
216
- - `/multi-agent:testflight-validation --resume` to re-run after fixes
217
- - `/multi-agent:fix-bug` when Gate 3 produced code-level findings
39
+ | Gate | What runs | Needs |
40
+ |---|---|---|
41
+ | **1. Static** | `ios_app_store_audit` (18 rules, real ITMS codes) | an `.xcarchive` |
42
+ | **2. Authoritative** | `ios_testflight_validate` → `altool --validate-app` | an `.ipa` + credentials |
43
+ | **3. Policy** | `app-store-review` skill vs repo source | repo checkout |
218
44
 
219
- Never upload, never bump the build number, never commit.
45
+ A skipped gate is never folded into the pass count, and the run never uploads. The
46
+ report lands in `~/StoreChecks/ios-<repo>-<branch>-<timestamp>/report.md`.
@@ -35,12 +35,36 @@ For Android: use `mcp__dev-toolkit__android_*` tools
35
35
  /multi-agent test "accessibility" -> accessibility audit (visual + MCP audit tool)
36
36
  /multi-agent test "dynamic type" -> large text test
37
37
  /multi-agent test "screenshot tr" -> locale screenshots
38
- /multi-agent test "store-ready" -> full audit: visual + accessibility + archive/APK compliance
39
- /multi-agent test "biometric" -> Face ID / Touch ID flow test
40
- /multi-agent test "performance" -> launch time + scroll performance
38
+ /multi-agent test "store-ready" -> hands off to /multi-agent:store-ready (package validation)
41
39
  /sim-test -> standalone (same thing)
42
40
  ```
43
41
 
42
+ Not offered, deliberately - `"biometric"` and `"performance"` were listed here with
43
+ no implementation section, so reaching one fell through to the general sweep and got
44
+ reported as the scenario asked for. Neither can be implemented symmetrically today:
45
+ biometric has an iOS tool (`ios_biometric`) and no Android counterpart, and launch
46
+ timing has `android_launch_time` and no iOS counterpart. Platform is auto-detected,
47
+ so either one would work on one platform and silently do nothing on the other. They
48
+ come back when the missing side exists, not before - do not re-add the rows to make
49
+ the list look complete.
50
+
51
+ Four scenarios also have a fixed-scenario command that pins the tag, so it does not
52
+ have to be typed or quoted. They arrive here with the scenario already resolved -
53
+ treat them as identical to the quoted form:
54
+
55
+ ```
56
+ /multi-agent:test-dark-mode -> scenario "dark mode"
57
+ /multi-agent:test-accessibility -> scenario "accessibility"
58
+ /multi-agent:test-dynamic-type -> scenario "dynamic type"
59
+ /multi-agent:test-screenshots [locale] -> scenario "screenshot <locale>", locale defaults to tr
60
+ ```
61
+
62
+ `store-ready` is not one of them, and does not belong to this file at all: it
63
+ validates a built package on either platform and lives at
64
+ `/multi-agent:store-ready`. The `"store-ready"` tag is kept as a hand-off so an
65
+ existing invocation still lands somewhere correct. `/multi-agent:testflight-validation`
66
+ is the iOS-pinned alias of that same command.
67
+
44
68
  ## Flow
45
69
 
46
70
  ### Step 1 - Device & App Discovery
@@ -121,43 +145,28 @@ Call: ios_set_locale(language: "<lang>", bundle_id: "...")
121
145
 
122
146
  **"store-ready":**
123
147
 
124
- Runs visual + accessibility pass AND dispatches the platform-matching compliance skill so the build is pre-validated against Apple / Google store requirements before submission.
148
+ Not implemented here. `store-ready` validates a built **package**, which is a
149
+ different job from driving a running app, and it is owned by one command on both
150
+ platforms:
125
151
 
126
152
  ```
127
- Platform detection:
128
- cwd contains .xcodeproj OR Package.swift -> iOS
129
- cwd contains build.gradle OR build.gradle.kts -> Android
130
- Otherwise -> error: "store-ready needs an iOS or Android project"
131
-
132
- Step A - run the standard visual + accessibility sweep (light + dark + large text).
133
-
134
- Step B - dispatch compliance skill:
135
- iOS:
136
- Load $HOME/.claude/skills/apple-archive-compliance/SKILL.md
137
- Prereq: @mmerterden/dev-toolkit-mcp ≥ v2.9.0 (provides ios_app_store_audit MCP tool)
138
- - installable via npm; standalone ArchiveGuard binary deprecated in v8.4.0
139
- Artifact: pick newest .xcarchive under ~/Library/Developer/Xcode/Archives/**
140
- OR accept explicit path argument after "store-ready"
141
- Invoke: mcp__dev-toolkit__ios_app_store_audit({ archive_path: <path>, rules: "all" })
142
- OR (Node fallback): node -e "import('@mmerterden/dev-toolkit-mcp/tools/ios-app-store-audit/index.js').then(m => m.runAudit({ archivePath: '<path>', rules: 'all' }).then(r => console.log(JSON.stringify(r))))"
143
- Humanize via --lang en (promptLanguage is locked to "en"; pass --lang=tr explicitly to opt into Turkish)
144
- Android:
145
- Load $HOME/.claude/skills/google-play-compliance/SKILL.md
146
- Prereq: bundletool, aapt2, apksigner in PATH (Android SDK build-tools)
147
- Artifact: newest .aab under **/build/outputs/bundle/**/*.aab
148
- OR explicit path argument
149
- Invoke: bundletool validate + dump manifest, aapt2 dump badging,
150
- apksigner verify, ABI/native scan
151
- Merge findings with UI-hunter findings -> single report, severity-grouped.
152
-
153
- Step C - humanize via `--lang en` by default (promptLanguage is locked to "en"; pass --lang=tr explicitly to opt into Turkish), group by severity
154
- (error = blocker, warning = risk, info = hygiene), link each finding to the
155
- Apple ITMS / Play policy reference supplied by the skill catalog.
156
-
157
- Step D - offer `/multi-agent:channels` follow-up so findings can land in
158
- Jira / Confluence / Wiki / PR body via the normal Phase 7 machinery.
153
+ $HOME/.claude/commands/multi-agent/store-ready/SKILL.md
159
154
  ```
160
155
 
156
+ Read that file and follow it. Pass through any artifact path given after
157
+ `store-ready` as its `--archive=` / `--ipa=` / `--aab=` / `--apk=` input.
158
+
159
+ Its Step A is the visual + accessibility sweep in this file - it calls back here
160
+ for the running-app half, then runs three gates per platform that nothing in this
161
+ file can do: the static package audit, the store's own validator, and a policy
162
+ review against repo source. It merges both halves into one severity-grouped report
163
+ and offers the `/multi-agent:channels` follow-up.
164
+
165
+ The archive-compliance audit used to be duplicated here, invoking
166
+ `ios_app_store_audit` with exactly the arguments the store-ready command's Gate 1
167
+ uses. Two copies of one call is how the iOS path grew a second door with no Gate 2,
168
+ no Gate 3 and no Android parity, so the copy is gone rather than kept in sync.
169
+
161
170
  **No argument (full test):**
162
171
 
163
172
  - Run light mode -> all screens
@@ -6,14 +6,15 @@
6
6
 
7
7
  ---
8
8
 
9
- ## 1. Command Inventory (44 commands)
9
+ ## 1. Command Inventory (49 commands)
10
10
 
11
11
  ```
12
12
  analysis, analysis-resolve, autopilot, build-optimize, channels, create-jira, design-check, dev,
13
13
  dev-autopilot, dev-local, dev-local-autopilot, diff-explain, forget, garbage-collect,
14
14
  help, ios-coding-standard, issue, jira, kill, language, local,
15
15
  local-autopilot, log, manual-test, prune-logs, purge, refactor, resume, review, review-issue, review-jira,
16
- routines, save, scan, search, setup, ship, stack, status, sync, test, testflight-validation, uninstall, update
16
+ routines, save, scan, search, setup, ship, stack, status, store-ready, sync, test, test-accessibility,
17
+ test-dark-mode, test-dynamic-type, test-screenshots, testflight-validation, uninstall, update
17
18
  ```
18
19
 
19
20
  Categories:
@@ -56,7 +56,7 @@ Knowledge files grow over time. Maintenance rules:
56
56
  - 90-day-old entries are considered "stale" - not used without verification
57
57
  - Stale check: orchestrator checks file mtime during Phase 1 knowledge injection
58
58
  - Stale entries are added to the prompt with a "STALE - verify before relying" tag
59
- - `clear-logs` does not touch knowledge - only deletes logs and state
59
+ - `prune-logs` does not touch knowledge - only deletes logs and state
60
60
  - `purge` does not touch knowledge either - separate command: `/multi-agent clear-knowledge {project}`
61
61
 
62
62
  ---
@@ -406,7 +406,7 @@ git -C $PROJECT_ROOT config user.email "{identity.email}"
406
406
 
407
407
  `worktreePath` = `$PROJECT_ROOT`, `localMode` = `true`.
408
408
 
409
- **If normal mode** (worktree - default): 2. Worktree path: Jira → `.worktrees/{jiraId}/`, GitHub → `.worktrees/GH{issueNo}/`, free-text → `.worktrees/task-{shortId}/` 3. **Heal stale admin state first** (see "Worktree stale-lock heal" below) and **apply the residue guard** (see "Worktree residue guard" below), then `git -C $PROJECT_ROOT worktree add {path} -b {branch} origin/{baseBranch}` (if exists: enter, pull) 4. Set identity: `git -C {worktree-path} config user.name/email` 5. Create log dir + `agent-log.md` + `agent-state.json`:
409
+ **If normal mode** (worktree - default): 2. Worktree path: Jira → `.worktrees/{jiraId}/`, GitHub → `.worktrees/GH{issueNo}/`, free-text → `.worktrees/task-{shortId}/` 3. **Heal stale admin state first** (see "Worktree stale-lock heal" below) and **apply the residue guard** (see "Worktree residue guard" below), then `git -C $PROJECT_ROOT worktree add {path} -b {branch} origin/{baseBranch}` (if exists: enter, pull) 4. Set identity: `git -C {worktree-path} config user.name/email` 5. Create log dir + `agent-log.md` + `agent-state.json` at `$HOME/.claude/logs/multi-agent/{project}/{task-id}/`, never inside the worktree:
410
410
 
411
411
  **Worktree stale-lock heal (required before every `worktree add`):** a run killed mid-`worktree add` (OOM, SIGTERM, disk full) leaves a locked or broken admin entry under `.git/worktrees/{id}/`, so the retry fails with `fatal: '<path>' already exists`. Always run the heal first - it is a no-op on a clean repo:
412
412
 
@@ -184,7 +184,7 @@
184
184
  },
185
185
  "appstore_connect_key_id": {
186
186
  "type": ["string", "null"],
187
- "description": "App Store Connect API key ID. Tier 1 of the App Store Connect access chain, used by /multi-agent:testflight-validation Gate 2. An identifier rather than a secret; mapped anyway so every credential is read through the same layer. Creating an API key needs an Admin or App Manager role, which is why Tier 2 exists."
187
+ "description": "App Store Connect API key ID. Tier 1 of the App Store Connect access chain, used by /multi-agent:store-ready Gate 2 (and its iOS alias /multi-agent:testflight-validation). An identifier rather than a secret; mapped anyway so every credential is read through the same layer. Creating an API key needs an Admin or App Manager role, which is why Tier 2 exists."
188
188
  },
189
189
  "appstore_connect_issuer_id": {
190
190
  "type": ["string", "null"],