deepseek-foreman 0.3.0 → 0.3.1

This diff represents the content of publicly available package versions that have been released to one of the supported registries. The information contained in this diff is provided for informational purposes only and reflects changes between package versions as they appear in their respective public registries.
Files changed (3) hide show
  1. package/README.en.md +10 -0
  2. package/README.md +10 -0
  3. package/package.json +1 -1
package/README.en.md CHANGED
@@ -176,6 +176,16 @@ The three pieces cover one segment each and add up:
176
176
 
177
177
  **Honest accounting**: the overall numbers from the 19 tickets are in [Field data](#field-data) above; per-plugin isolated quantification **has had no A/B control, so no numbers are given** — only mechanisms: build dominates token spend → cheap workers; dialogue and judgement stay terse → persona; quality holds → cross-vendor review + item-by-item verification. The three mechanisms each own one segment, and they stack.
178
178
 
179
+ ## Ticket tiers (effort scaling)
180
+
181
+ | Tier | Flow | Dispatch effort |
182
+ |---|---|---|
183
+ | `trivial` (≤10 lines/docs) | skip cross-vendor review, Lead accepts | low |
184
+ | `normal` (default) | full loop + cheap cross-vendor review | per workers.md |
185
+ | `critical` (data/release/security) | double review + physical read-only | high; escalate after 2 failures |
186
+
187
+ Every receipt logs a cost ledger (build token/effort/wall-clock); budget caps halt unattended runs. Model-role logic and measured data: [Quantified metrics](docs/metrics-2026-10-01.md).
188
+
179
189
  ## Status
180
190
 
181
191
  Installed into dsh desktop 0.2.0-rc.2 (Intel iMac) and verified live: cold start is normal, `pick_route` is callable by the model, the vision constraint really blocks and returns a fallback; the final `subagent_readonly` state is live-verified too (read-only tool set enforced, delegation tools blocked).
package/README.md CHANGED
@@ -174,6 +174,16 @@ bundle 的 `cordis.patch.yml` 里**新增插件行要包在 `insert:` 下**:
174
174
 
175
175
  **口径(诚实)**:19 张工单的整体数据见上文[实测数据](#实测数据);逐插件的孤立量化贡献**没做 A/B 对照,不给数字**,只讲机制——token 大头在实现 → 便宜工人;对话与判断从简 → persona 精简;质量不降 → 异族审查 + 逐条核实。三个机制各管一段,叠加生效。
176
176
 
177
+ ## 工单分级(effort scaling)
178
+
179
+ | 级别 | 走什么流程 | 派单 effort |
180
+ |---|---|---|
181
+ | `trivial`(≤10 行/纯文档) | 跳过异族审查,Lead 验收收尾 | low |
182
+ | `normal`(默认) | 完整闭环 + 异族便宜审查 | 按 workers.md |
183
+ | `critical`(数据/发布/安全) | 双审查 + 物理只读审查 | high;失败 2 次升档重派 |
184
+
185
+ 成本台账随每张回执记录(施工 token/effort/wall-clock),`_receipts/progress.md` 周汇总;无人托管有预算闸(到限即停)。分工逻辑与实测数据见 [可量化测试数据](docs/metrics-2026-10-01.md)。
186
+
177
187
  ## 状态
178
188
 
179
189
  已装进 dsh 桌面版 0.2.0-rc.2(Intel iMac)并 live 验证:冷启动正常,`pick_route` 可被模型调用,视觉约束会真的拦截并给 fallback;`subagent_readonly` 最终态同样 live 实测通过(只读工具集生效、委派工具被拦死)。
package/package.json CHANGED
@@ -1,6 +1,6 @@
1
1
  {
2
2
  "name": "deepseek-foreman",
3
- "version": "0.3.0",
3
+ "version": "0.3.1",
4
4
  "description": "Foreman for DeepSeek Harness: your best model leads, cheaper models build, a rival vendor reviews. Role-to-route adjudication with peak-window, vision, output-size and cross-vendor-review constraints.",
5
5
  "type": "module",
6
6
  "main": "lib/index.js",