@gkorepanov/ccodex 0.4.15 → 0.4.17

This diff represents the content of publicly available package versions that have been released to one of the supported registries. The information contained in this diff is provided for informational purposes only and reflects changes between package versions as they appear in their respective public registries.
@@ -0,0 +1,28 @@
1
+ #!/bin/sh
2
+ # Install the Claude -> Codex delegation stack into Claude Code:
3
+ # codex-wrapper agent, skills (workforce, ...), and the codex MCP server.
4
+ set -eu
5
+
6
+ ROOT=$(CDPATH= cd -- "$(dirname -- "$0")/.." && pwd)
7
+ CLAUDE_DIR=${CLAUDE_DIR:-"$HOME/.claude"}
8
+
9
+ mkdir -p "$CLAUDE_DIR/agents" "$CLAUDE_DIR/skills"
10
+
11
+ cp "$ROOT/agents/codex-wrapper.md" "$CLAUDE_DIR/agents/codex-wrapper.md"
12
+ echo "installed agent: codex-wrapper -> $CLAUDE_DIR/agents/codex-wrapper.md"
13
+
14
+ for skill in "$ROOT"/skills/*/; do
15
+ name=$(basename "$skill")
16
+ rm -rf "$CLAUDE_DIR/skills/$name"
17
+ cp -R "$skill" "$CLAUDE_DIR/skills/$name"
18
+ echo "installed skill: $name -> $CLAUDE_DIR/skills/$name"
19
+ done
20
+
21
+ if ! command -v claude >/dev/null 2>&1; then
22
+ echo "codex MCP server: skipped (claude CLI not found)" >&2
23
+ elif claude mcp get codex >/dev/null 2>&1; then
24
+ echo "codex MCP server: already registered"
25
+ else
26
+ claude mcp add --scope user codex -- codex mcp-server
27
+ echo "codex MCP server: registered (user scope)"
28
+ fi
@@ -0,0 +1,35 @@
1
+ ---
2
+ name: workforce
3
+ description: "Delegation policy and model rankings for subagents: which model (gpt-5.6-sol / opus-5 / fable-5) to pick for which task, how to be an orchestrator instead of burning your own context, and how to run gpt-5.6-sol via codex-wrapper. Use BEFORE delegating any work to subagents, spawning agents/workflows, or choosing a model for a subtask."
4
+ ---
5
+
6
+ # Subagents and token usage
7
+
8
+ Be very careful with token usage: be an orchestrator and manager over subagents. Do not dive into huge repositories, read whole huge docs/files, or write routine code/configs/migrations/plots yourself — delegate all token-heavy dirty work to subagents. Spend your own context only on planning, architecture, research taste, code design, reading tough places where less capable agents struggle, and reviewing (or writing) production/good code.
9
+ Especially avoid (!!!) baby-sitting long runs (e.g. training or feature collection) yourself — let gpt-5.6-sol do it for you.
10
+ One more time: NEVER spend your own time/tokens on manual labour such as writing a well-scoped, verifiable, non-production module with clearly defined inputs and outputs.
11
+
12
+ ## Rankings
13
+
14
+ Axes 0–10: INT = intelligence (how hard a well-defined problem it handles unsupervised), RT = research taste, PCQ = production code quality, IF = exact instruction following, ATD = attention to detail. Cost = what I actually pay (OpenAI limits are generous).
15
+
16
+ | model | cost | INT | RT | PCQ | IF | ATD |
17
+ |--------------------|------|-----|----|-----|----|-----|
18
+ | gpt-5.6-sol (high) | $ | 8 | 5 | 4 | 10 | 10 |
19
+ | opus-5 (high) | $$ | 8 | 8 | 6 | 8 | 7 |
20
+ | fable-5 (high) | $$$$ | 9 | 9 | 9 | 8 | 8 |
21
+
22
+ How to apply:
23
+ - Pick the cheapest model whose scores meet the task's bar; when axes conflict for anything that ships, INT & RT >> cost.
24
+ - NEVER delegate open-ended / not-well-specified tasks that require independent research and research taste to gpt-5.6-sol (e.g. "research why strategy X has a quality drawdown"). Do them yourself or delegate to a fable-5/opus-5 subagent. Delegate to gpt-5.6-sol only engineering or well-specified tasks where the spec fully defines the result: "implement a strategy with such-and-such logic" — ok; "build such-and-such plot with such-and-such math" — ok; open-ended investigation — not ok.
25
+ - Though: you can always run gpt-5.6-sol on any research in parallel — never trust its _research_ conclusions blindly, but it can spot details that e.g. fable-5 might miss.
26
+ - If a cheaper model's output doesn't meet the bar, redo the work with a smarter model without asking me (model choice is yours; the autonomy rules still gate *what* work is allowed). Judge the output, not the price tag: escalating costs less than shipping mediocre work.
27
+ - Never trust agent conclusions blindly. When an agent claims "X is better than Y because <arguments>", the arguments must be backed by substantive, simple, easily explainable cases. If a subagent explains its conclusion in an overly convoluted way — a pile of numbers and plots instead of a clear story — treat the conclusion as unreliable and send another independent subagent to re-verify the ESSENCE.
28
+ - Bulk/mechanical work (explore current codebase, clear-spec implementation, straightforward data analysis, migrations): gpt-5.6-sol — its IF/ATD are top, but PCQ 4 means never ask it to write production code.
29
+ - Anything user-facing (UI, copywriting, API design) needs RT ≥ 7.
30
+ - Reviews of plans/implementations: fable-5. Optionally gpt-5.6-sol as an extra independent perspective.
31
+ - Never use models below the table (Haiku, bare Sonnet, etc.) for actual work.
32
+ - Do not use Explore subagent for codebase exploration, use gpt-5.6-sol instead.
33
+
34
+ ## Using gpt-5.6-sol
35
+ Claude models (opus-5, fable-5) run via the Agent/Workflow `model` parameter; that parameter only takes Claude models, so for gpt-5.6-sol use a `codex-wrapper` subagent (`model: 'sonnet', effort: 'low'`) which is instructed to communicate with gpt-5.6-sol and return the result.