deepseek-foreman 0.3.0 → 0.3.1
This diff represents the content of publicly available package versions that have been released to one of the supported registries. The information contained in this diff is provided for informational purposes only and reflects changes between package versions as they appear in their respective public registries.
- package/README.en.md +10 -0
- package/README.md +10 -0
- package/package.json +1 -1
package/README.en.md
CHANGED
|
@@ -176,6 +176,16 @@ The three pieces cover one segment each and add up:
|
|
|
176
176
|
|
|
177
177
|
**Honest accounting**: the overall numbers from the 19 tickets are in [Field data](#field-data) above; per-plugin isolated quantification **has had no A/B control, so no numbers are given** — only mechanisms: build dominates token spend → cheap workers; dialogue and judgement stay terse → persona; quality holds → cross-vendor review + item-by-item verification. The three mechanisms each own one segment, and they stack.
|
|
178
178
|
|
|
179
|
+
## Ticket tiers (effort scaling)
|
|
180
|
+
|
|
181
|
+
| Tier | Flow | Dispatch effort |
|
|
182
|
+
|---|---|---|
|
|
183
|
+
| `trivial` (≤10 lines/docs) | skip cross-vendor review, Lead accepts | low |
|
|
184
|
+
| `normal` (default) | full loop + cheap cross-vendor review | per workers.md |
|
|
185
|
+
| `critical` (data/release/security) | double review + physical read-only | high; escalate after 2 failures |
|
|
186
|
+
|
|
187
|
+
Every receipt logs a cost ledger (build token/effort/wall-clock); budget caps halt unattended runs. Model-role logic and measured data: [Quantified metrics](docs/metrics-2026-10-01.md).
|
|
188
|
+
|
|
179
189
|
## Status
|
|
180
190
|
|
|
181
191
|
Installed into dsh desktop 0.2.0-rc.2 (Intel iMac) and verified live: cold start is normal, `pick_route` is callable by the model, the vision constraint really blocks and returns a fallback; the final `subagent_readonly` state is live-verified too (read-only tool set enforced, delegation tools blocked).
|
package/README.md
CHANGED
|
@@ -174,6 +174,16 @@ bundle 的 `cordis.patch.yml` 里**新增插件行要包在 `insert:` 下**:
|
|
|
174
174
|
|
|
175
175
|
**口径(诚实)**:19 张工单的整体数据见上文[实测数据](#实测数据);逐插件的孤立量化贡献**没做 A/B 对照,不给数字**,只讲机制——token 大头在实现 → 便宜工人;对话与判断从简 → persona 精简;质量不降 → 异族审查 + 逐条核实。三个机制各管一段,叠加生效。
|
|
176
176
|
|
|
177
|
+
## 工单分级(effort scaling)
|
|
178
|
+
|
|
179
|
+
| 级别 | 走什么流程 | 派单 effort |
|
|
180
|
+
|---|---|---|
|
|
181
|
+
| `trivial`(≤10 行/纯文档) | 跳过异族审查,Lead 验收收尾 | low |
|
|
182
|
+
| `normal`(默认) | 完整闭环 + 异族便宜审查 | 按 workers.md |
|
|
183
|
+
| `critical`(数据/发布/安全) | 双审查 + 物理只读审查 | high;失败 2 次升档重派 |
|
|
184
|
+
|
|
185
|
+
成本台账随每张回执记录(施工 token/effort/wall-clock),`_receipts/progress.md` 周汇总;无人托管有预算闸(到限即停)。分工逻辑与实测数据见 [可量化测试数据](docs/metrics-2026-10-01.md)。
|
|
186
|
+
|
|
177
187
|
## 状态
|
|
178
188
|
|
|
179
189
|
已装进 dsh 桌面版 0.2.0-rc.2(Intel iMac)并 live 验证:冷启动正常,`pick_route` 可被模型调用,视觉约束会真的拦截并给 fallback;`subagent_readonly` 最终态同样 live 实测通过(只读工具集生效、委派工具被拦死)。
|
package/package.json
CHANGED
|
@@ -1,6 +1,6 @@
|
|
|
1
1
|
{
|
|
2
2
|
"name": "deepseek-foreman",
|
|
3
|
-
"version": "0.3.
|
|
3
|
+
"version": "0.3.1",
|
|
4
4
|
"description": "Foreman for DeepSeek Harness: your best model leads, cheaper models build, a rival vendor reviews. Role-to-route adjudication with peak-window, vision, output-size and cross-vendor-review constraints.",
|
|
5
5
|
"type": "module",
|
|
6
6
|
"main": "lib/index.js",
|