dsh-aris-panel 0.1.0

This diff represents the content of publicly available package versions that have been released to one of the supported registries. The information contained in this diff is provided for informational purposes only and reflects changes between package versions as they appear in their respective public registries.
Files changed (442) hide show
  1. package/LICENSE +21 -0
  2. package/README.md +98 -0
  3. package/README_CN.md +87 -0
  4. package/dsh/checkout.patch.yml +38 -0
  5. package/dsh/client.js +634 -0
  6. package/dsh/cordis.patch.yml +44 -0
  7. package/dsh/index.mjs +76 -0
  8. package/dsh/run-status.mjs +182 -0
  9. package/dsh/scope-limits.mjs +50 -0
  10. package/dsh/workbench.mjs +291 -0
  11. package/mcp-servers/claude-review/README.md +93 -0
  12. package/mcp-servers/claude-review/run_with_claude_aws.sh +49 -0
  13. package/mcp-servers/claude-review/server.py +718 -0
  14. package/mcp-servers/codex-image2/README.md +65 -0
  15. package/mcp-servers/codex-image2/server.py +893 -0
  16. package/mcp-servers/feishu-bridge/requirements.txt +1 -0
  17. package/mcp-servers/feishu-bridge/server.py +240 -0
  18. package/mcp-servers/gemini-review/README.md +171 -0
  19. package/mcp-servers/gemini-review/server.py +1856 -0
  20. package/mcp-servers/llm-chat/requirements.txt +1 -0
  21. package/mcp-servers/llm-chat/server.py +664 -0
  22. package/mcp-servers/manual-review/README.md +133 -0
  23. package/mcp-servers/manual-review/server.py +910 -0
  24. package/mcp-servers/manual-review/ui.html +279 -0
  25. package/mcp-servers/minimax-chat/requirements.txt +1 -0
  26. package/mcp-servers/minimax-chat/server.py +381 -0
  27. package/package.json +51 -0
  28. package/skills/ablation-planner/SKILL.md +123 -0
  29. package/skills/alphaxiv/SKILL.md +196 -0
  30. package/skills/analyze-results/SKILL.md +46 -0
  31. package/skills/arxiv/SKILL.md +248 -0
  32. package/skills/auto-paper-improvement-loop/SKILL.md +651 -0
  33. package/skills/auto-review-loop/SKILL.md +1137 -0
  34. package/skills/auto-review-loop-llm/SKILL.md +259 -0
  35. package/skills/auto-review-loop-minimax/SKILL.md +302 -0
  36. package/skills/citation-audit/SKILL.md +502 -0
  37. package/skills/claims-drafting/SKILL.md +227 -0
  38. package/skills/comm-lit-review/SKILL.md +297 -0
  39. package/skills/deepxiv/SKILL.md +263 -0
  40. package/skills/dse-loop/SKILL.md +296 -0
  41. package/skills/embodiment-description/SKILL.md +129 -0
  42. package/skills/exa-search/SKILL.md +205 -0
  43. package/skills/experiment-audit/SKILL.md +311 -0
  44. package/skills/experiment-bridge/SKILL.md +376 -0
  45. package/skills/experiment-plan/SKILL.md +249 -0
  46. package/skills/experiment-queue/SKILL.md +431 -0
  47. package/skills/experiment-queue/scripts/build_manifest.py +142 -0
  48. package/skills/experiment-queue/scripts/queue_manager.py +433 -0
  49. package/skills/feishu-notify/SKILL.md +156 -0
  50. package/skills/figure-description/SKILL.md +138 -0
  51. package/skills/figure-spec/SKILL.md +262 -0
  52. package/skills/figure-spec/scripts/figure_renderer.py +799 -0
  53. package/skills/formula-derivation/SKILL.md +280 -0
  54. package/skills/gemini-search/SKILL.md +231 -0
  55. package/skills/grant-proposal/SKILL.md +698 -0
  56. package/skills/idea-creator/SKILL.md +542 -0
  57. package/skills/idea-discovery/SKILL.md +521 -0
  58. package/skills/idea-discovery-robot/SKILL.md +363 -0
  59. package/skills/integrity-forensics/SKILL.md +284 -0
  60. package/skills/interview-cheatsheet/SKILL.md +245 -0
  61. package/skills/invention-structuring/SKILL.md +188 -0
  62. package/skills/jurisdiction-format/SKILL.md +192 -0
  63. package/skills/kill-argument/SKILL.md +437 -0
  64. package/skills/mermaid-diagram/SKILL.md +419 -0
  65. package/skills/meta-apply/SKILL.md +141 -0
  66. package/skills/meta-optimize/SKILL.md +437 -0
  67. package/skills/monitor-experiment/SKILL.md +140 -0
  68. package/skills/novelty-check/SKILL.md +101 -0
  69. package/skills/openalex/SKILL.md +237 -0
  70. package/skills/overleaf-sync/SKILL.md +220 -0
  71. package/skills/paper-claim-audit/SKILL.md +348 -0
  72. package/skills/paper-compile/SKILL.md +266 -0
  73. package/skills/paper-figure/SKILL.md +312 -0
  74. package/skills/paper-illustration/SKILL.md +736 -0
  75. package/skills/paper-illustration-image2/SKILL.md +391 -0
  76. package/skills/paper-illustration-image2/scripts/paper_illustration_image2.py +255 -0
  77. package/skills/paper-plan/SKILL.md +386 -0
  78. package/skills/paper-poster/SKILL.md +19 -0
  79. package/skills/paper-poster-html/DESIGN_FINAL.md +176 -0
  80. package/skills/paper-poster-html/IMPLEMENTATION_CONVENTIONS.md +161 -0
  81. package/skills/paper-poster-html/LICENSES/posterly-MIT.txt +21 -0
  82. package/skills/paper-poster-html/NOTICE.md +57 -0
  83. package/skills/paper-poster-html/SKILL.md +323 -0
  84. package/skills/paper-poster-html/scripts/_posterly/__init__.py +0 -0
  85. package/skills/paper-poster-html/scripts/_posterly/canvas.py +200 -0
  86. package/skills/paper-poster-html/scripts/_posterly/measure.py +588 -0
  87. package/skills/paper-poster-html/scripts/_posterly/polish.py +498 -0
  88. package/skills/paper-poster-html/scripts/_posterly/preflight.py +489 -0
  89. package/skills/paper-poster-html/scripts/_posterly/render.py +215 -0
  90. package/skills/paper-poster-html/scripts/_posterly/textutil.py +16 -0
  91. package/skills/paper-poster-html/scripts/_posterly/verify_final.py +171 -0
  92. package/skills/paper-poster-html/scripts/asset_check.py +897 -0
  93. package/skills/paper-poster-html/scripts/extract_pdf_figures.py +666 -0
  94. package/skills/paper-poster-html/scripts/poster_check.py +251 -0
  95. package/skills/paper-poster-html/scripts/preprocess_figures.py +238 -0
  96. package/skills/paper-poster-html/scripts/render_preview.py +217 -0
  97. package/skills/paper-poster-html/scripts/run_gates.py +556 -0
  98. package/skills/paper-poster-html/scripts/style_check.py +1324 -0
  99. package/skills/paper-poster-html/templates/COMPONENTS.md +462 -0
  100. package/skills/paper-poster-html/templates/README.md +170 -0
  101. package/skills/paper-poster-html/templates/landscape_4col.html +1032 -0
  102. package/skills/paper-poster-html/templates/landscape_hero.html +1046 -0
  103. package/skills/paper-poster-html/templates/portrait_2col.html +947 -0
  104. package/skills/paper-poster-html/templates/tokens/acl.json +9 -0
  105. package/skills/paper-poster-html/templates/tokens/cvpr.json +9 -0
  106. package/skills/paper-poster-html/templates/tokens/generic.json +9 -0
  107. package/skills/paper-poster-html/templates/tokens/iclr.json +9 -0
  108. package/skills/paper-poster-html/templates/tokens/icml.json +9 -0
  109. package/skills/paper-poster-html/templates/tokens/neurips.json +9 -0
  110. package/skills/paper-slides/SKILL.md +635 -0
  111. package/skills/paper-talk/SKILL.md +381 -0
  112. package/skills/paper-write/SKILL.md +604 -0
  113. package/skills/paper-write/templates/IEEEtran.bst +2409 -0
  114. package/skills/paper-write/templates/IEEEtran.cls +6347 -0
  115. package/skills/paper-write/templates/iclr2026.tex +84 -0
  116. package/skills/paper-write/templates/icml2025.tex +87 -0
  117. package/skills/paper-write/templates/ieee_conference.tex +89 -0
  118. package/skills/paper-write/templates/ieee_journal.tex +93 -0
  119. package/skills/paper-write/templates/math_commands.tex +48 -0
  120. package/skills/paper-write/templates/neurips2025.tex +80 -0
  121. package/skills/paper-writing/SKILL.md +916 -0
  122. package/skills/patent-novelty-check/SKILL.md +153 -0
  123. package/skills/patent-pipeline/SKILL.md +344 -0
  124. package/skills/patent-review/SKILL.md +203 -0
  125. package/skills/pixel-art/SKILL.md +137 -0
  126. package/skills/prior-art-search/SKILL.md +146 -0
  127. package/skills/proof-checker/SKILL.md +866 -0
  128. package/skills/proof-orchestrator/NOTICE.md +24 -0
  129. package/skills/proof-orchestrator/SKILL.md +254 -0
  130. package/skills/proof-orchestrator/references/audit-output-contract.md +126 -0
  131. package/skills/proof-orchestrator/references/deepseek-routing.md +74 -0
  132. package/skills/proof-orchestrator/references/dispatch-prompts.md +227 -0
  133. package/skills/proof-orchestrator/references/notation-audit.md +135 -0
  134. package/skills/proof-orchestrator/references/proof-audit-rubric.md +70 -0
  135. package/skills/proof-orchestrator/references/stress-tests.md +38 -0
  136. package/skills/proof-writer/SKILL.md +223 -0
  137. package/skills/qzcli/SKILL.md +324 -0
  138. package/skills/rebuttal/SKILL.md +376 -0
  139. package/skills/render-html/SKILL.md +316 -0
  140. package/skills/render-html/scripts/render_html.py +1006 -0
  141. package/skills/render-html/scripts/templates/academic.html +703 -0
  142. package/skills/render-html/scripts/templates/dashboard.html +333 -0
  143. package/skills/research-lit/SKILL.md +756 -0
  144. package/skills/research-pipeline/SKILL.md +384 -0
  145. package/skills/research-refine/SKILL.md +770 -0
  146. package/skills/research-refine-pipeline/SKILL.md +186 -0
  147. package/skills/research-review/SKILL.md +198 -0
  148. package/skills/research-wiki/SKILL.md +461 -0
  149. package/skills/resubmit-pipeline/SKILL.md +447 -0
  150. package/skills/result-to-claim/SKILL.md +311 -0
  151. package/skills/run-experiment/SKILL.md +313 -0
  152. package/skills/semantic-scholar/SKILL.md +236 -0
  153. package/skills/serverless-modal/SKILL.md +335 -0
  154. package/skills/shared-references/acceptance-gate.md +324 -0
  155. package/skills/shared-references/assurance-contract.md +248 -0
  156. package/skills/shared-references/capture-antipatterns.md +78 -0
  157. package/skills/shared-references/citation-discipline.md +583 -0
  158. package/skills/shared-references/compute-env-contract.md +163 -0
  159. package/skills/shared-references/effort-contract.md +183 -0
  160. package/skills/shared-references/evidence-precheck.md +65 -0
  161. package/skills/shared-references/experiment-integrity.md +49 -0
  162. package/skills/shared-references/external-cadence.md +326 -0
  163. package/skills/shared-references/fan-out-pattern.md +366 -0
  164. package/skills/shared-references/injection-hygiene.md +127 -0
  165. package/skills/shared-references/integration-contract.md +461 -0
  166. package/skills/shared-references/output-composition.md +93 -0
  167. package/skills/shared-references/output-language.md +45 -0
  168. package/skills/shared-references/output-manifest.md +49 -0
  169. package/skills/shared-references/output-versioning.md +111 -0
  170. package/skills/shared-references/patent-format-cn.md +199 -0
  171. package/skills/shared-references/patent-format-ep.md +173 -0
  172. package/skills/shared-references/patent-format-us.md +161 -0
  173. package/skills/shared-references/patent-writing-principles.md +197 -0
  174. package/skills/shared-references/prior-art-databases.md +141 -0
  175. package/skills/shared-references/resumable-runs.md +109 -0
  176. package/skills/shared-references/review-scope-limits.md +81 -0
  177. package/skills/shared-references/review-tracing.md +391 -0
  178. package/skills/shared-references/reviewer-independence.md +79 -0
  179. package/skills/shared-references/reviewer-routing.md +852 -0
  180. package/skills/shared-references/skill-governance.md +104 -0
  181. package/skills/shared-references/taste-calibration.md +85 -0
  182. package/skills/shared-references/venue-checklists.md +114 -0
  183. package/skills/shared-references/wiki-helper-resolution.md +134 -0
  184. package/skills/shared-references/writing-principles.md +525 -0
  185. package/skills/skills-codex/README.md +102 -0
  186. package/skills/skills-codex/README_CN.md +100 -0
  187. package/skills/skills-codex/ablation-planner/SKILL.md +126 -0
  188. package/skills/skills-codex/alphaxiv/SKILL.md +186 -0
  189. package/skills/skills-codex/analyze-results/SKILL.md +45 -0
  190. package/skills/skills-codex/arxiv/SKILL.md +210 -0
  191. package/skills/skills-codex/auto-paper-improvement-loop/SKILL.md +574 -0
  192. package/skills/skills-codex/auto-review-loop/SKILL.md +500 -0
  193. package/skills/skills-codex/auto-review-loop-llm/SKILL.md +247 -0
  194. package/skills/skills-codex/auto-review-loop-minimax/SKILL.md +290 -0
  195. package/skills/skills-codex/citation-audit/SKILL.md +504 -0
  196. package/skills/skills-codex/claims-drafting/SKILL.md +239 -0
  197. package/skills/skills-codex/comm-lit-review/SKILL.md +299 -0
  198. package/skills/skills-codex/comm-lit-review/references/domain-taxonomy.md +57 -0
  199. package/skills/skills-codex/comm-lit-review/references/output-template.md +37 -0
  200. package/skills/skills-codex/comm-lit-review/references/source-policy.md +99 -0
  201. package/skills/skills-codex/comm-lit-review/references/venue-tiering.md +112 -0
  202. package/skills/skills-codex/deepxiv/SKILL.md +142 -0
  203. package/skills/skills-codex/dse-loop/SKILL.md +285 -0
  204. package/skills/skills-codex/embodiment-description/SKILL.md +129 -0
  205. package/skills/skills-codex/exa-search/SKILL.md +192 -0
  206. package/skills/skills-codex/experiment-audit/SKILL.md +286 -0
  207. package/skills/skills-codex/experiment-bridge/SKILL.md +356 -0
  208. package/skills/skills-codex/experiment-plan/SKILL.md +249 -0
  209. package/skills/skills-codex/experiment-queue/SKILL.md +401 -0
  210. package/skills/skills-codex/feishu-notify/SKILL.md +155 -0
  211. package/skills/skills-codex/figure-description/SKILL.md +138 -0
  212. package/skills/skills-codex/figure-spec/SKILL.md +252 -0
  213. package/skills/skills-codex/formula-derivation/SKILL.md +280 -0
  214. package/skills/skills-codex/gemini-search/SKILL.md +205 -0
  215. package/skills/skills-codex/grant-proposal/SKILL.md +626 -0
  216. package/skills/skills-codex/idea-creator/SKILL.md +405 -0
  217. package/skills/skills-codex/idea-discovery/SKILL.md +475 -0
  218. package/skills/skills-codex/idea-discovery-robot/SKILL.md +362 -0
  219. package/skills/skills-codex/integrity-forensics/SKILL.md +106 -0
  220. package/skills/skills-codex/interview-cheatsheet/SKILL.md +245 -0
  221. package/skills/skills-codex/invention-structuring/SKILL.md +188 -0
  222. package/skills/skills-codex/jurisdiction-format/SKILL.md +192 -0
  223. package/skills/skills-codex/kill-argument/SKILL.md +403 -0
  224. package/skills/skills-codex/mermaid-diagram/SKILL.md +379 -0
  225. package/skills/skills-codex/meta-apply/SKILL.md +154 -0
  226. package/skills/skills-codex/meta-optimize/SKILL.md +348 -0
  227. package/skills/skills-codex/monitor-experiment/SKILL.md +98 -0
  228. package/skills/skills-codex/novelty-check/SKILL.md +89 -0
  229. package/skills/skills-codex/openalex/SKILL.md +228 -0
  230. package/skills/skills-codex/overleaf-sync/SKILL.md +220 -0
  231. package/skills/skills-codex/paper-claim-audit/SKILL.md +350 -0
  232. package/skills/skills-codex/paper-compile/SKILL.md +253 -0
  233. package/skills/skills-codex/paper-figure/SKILL.md +311 -0
  234. package/skills/skills-codex/paper-illustration/SKILL.md +690 -0
  235. package/skills/skills-codex/paper-illustration-image2/SKILL.md +383 -0
  236. package/skills/skills-codex/paper-illustration-image2/scripts/paper_illustration_image2.py +255 -0
  237. package/skills/skills-codex/paper-plan/SKILL.md +278 -0
  238. package/skills/skills-codex/paper-poster/SKILL.md +19 -0
  239. package/skills/skills-codex/paper-poster-html/SKILL.md +377 -0
  240. package/skills/skills-codex/paper-slides/SKILL.md +571 -0
  241. package/skills/skills-codex/paper-talk/SKILL.md +381 -0
  242. package/skills/skills-codex/paper-write/SKILL.md +411 -0
  243. package/skills/skills-codex/paper-write/templates/IEEEtran.bst +2409 -0
  244. package/skills/skills-codex/paper-write/templates/IEEEtran.cls +6347 -0
  245. package/skills/skills-codex/paper-write/templates/aaai2026.bst +1493 -0
  246. package/skills/skills-codex/paper-write/templates/aaai2026.sty +315 -0
  247. package/skills/skills-codex/paper-write/templates/aaai2026.tex +952 -0
  248. package/skills/skills-codex/paper-write/templates/acl.sty +312 -0
  249. package/skills/skills-codex/paper-write/templates/acl2026.tex +377 -0
  250. package/skills/skills-codex/paper-write/templates/acl_natbib.bst +1940 -0
  251. package/skills/skills-codex/paper-write/templates/acm.bst +3081 -0
  252. package/skills/skills-codex/paper-write/templates/acm_mm2026.tex +204 -0
  253. package/skills/skills-codex/paper-write/templates/acmart.cls +3520 -0
  254. package/skills/skills-codex/paper-write/templates/cvpr.bst +1448 -0
  255. package/skills/skills-codex/paper-write/templates/cvpr.sty +508 -0
  256. package/skills/skills-codex/paper-write/templates/cvpr2026.tex +63 -0
  257. package/skills/skills-codex/paper-write/templates/iclr2026.tex +84 -0
  258. package/skills/skills-codex/paper-write/templates/iclr2026_conference.bst +1440 -0
  259. package/skills/skills-codex/paper-write/templates/iclr2026_conference.sty +246 -0
  260. package/skills/skills-codex/paper-write/templates/icml2026.sty +767 -0
  261. package/skills/skills-codex/paper-write/templates/icml2026.tex +662 -0
  262. package/skills/skills-codex/paper-write/templates/ieee_conference.tex +89 -0
  263. package/skills/skills-codex/paper-write/templates/ieee_journal.tex +93 -0
  264. package/skills/skills-codex/paper-write/templates/math_commands.tex +48 -0
  265. package/skills/skills-codex/paper-write/templates/neurips2026.tex +493 -0
  266. package/skills/skills-codex/paper-write/templates/neurips_2026.sty +437 -0
  267. package/skills/skills-codex/paper-writing/SKILL.md +731 -0
  268. package/skills/skills-codex/patent-novelty-check/SKILL.md +153 -0
  269. package/skills/skills-codex/patent-pipeline/SKILL.md +344 -0
  270. package/skills/skills-codex/patent-review/SKILL.md +202 -0
  271. package/skills/skills-codex/pixel-art/SKILL.md +139 -0
  272. package/skills/skills-codex/prior-art-search/SKILL.md +146 -0
  273. package/skills/skills-codex/proof-checker/SKILL.md +554 -0
  274. package/skills/skills-codex/proof-orchestrator/SKILL.md +260 -0
  275. package/skills/skills-codex/proof-orchestrator/references/audit-output-contract.md +126 -0
  276. package/skills/skills-codex/proof-orchestrator/references/deepseek-routing.md +76 -0
  277. package/skills/skills-codex/proof-orchestrator/references/dispatch-prompts.md +227 -0
  278. package/skills/skills-codex/proof-orchestrator/references/notation-audit.md +135 -0
  279. package/skills/skills-codex/proof-orchestrator/references/proof-audit-rubric.md +70 -0
  280. package/skills/skills-codex/proof-orchestrator/references/stress-tests.md +38 -0
  281. package/skills/skills-codex/proof-writer/SKILL.md +222 -0
  282. package/skills/skills-codex/qzcli/SKILL.md +324 -0
  283. package/skills/skills-codex/rebuttal/SKILL.md +305 -0
  284. package/skills/skills-codex/render-html/SKILL.md +305 -0
  285. package/skills/skills-codex/render-html/scripts/__pycache__/render_html.cpython-314.pyc +0 -0
  286. package/skills/skills-codex/render-html/scripts/render_html.py +909 -0
  287. package/skills/skills-codex/render-html/scripts/templates/academic.html +342 -0
  288. package/skills/skills-codex/render-html/scripts/templates/dashboard.html +333 -0
  289. package/skills/skills-codex/research-lit/SKILL.md +464 -0
  290. package/skills/skills-codex/research-pipeline/SKILL.md +340 -0
  291. package/skills/skills-codex/research-refine/SKILL.md +721 -0
  292. package/skills/skills-codex/research-refine-pipeline/SKILL.md +186 -0
  293. package/skills/skills-codex/research-review/SKILL.md +135 -0
  294. package/skills/skills-codex/research-wiki/SKILL.md +421 -0
  295. package/skills/skills-codex/resubmit-pipeline/SKILL.md +444 -0
  296. package/skills/skills-codex/result-to-claim/SKILL.md +246 -0
  297. package/skills/skills-codex/run-experiment/SKILL.md +236 -0
  298. package/skills/skills-codex/semantic-scholar/SKILL.md +219 -0
  299. package/skills/skills-codex/serverless-modal/SKILL.md +335 -0
  300. package/skills/skills-codex/shared-references/acceptance-gate.md +336 -0
  301. package/skills/skills-codex/shared-references/assurance-contract.md +139 -0
  302. package/skills/skills-codex/shared-references/capture-antipatterns.md +84 -0
  303. package/skills/skills-codex/shared-references/citation-discipline.md +452 -0
  304. package/skills/skills-codex/shared-references/compute-env-contract.md +163 -0
  305. package/skills/skills-codex/shared-references/effort-contract.md +143 -0
  306. package/skills/skills-codex/shared-references/evidence-precheck.md +73 -0
  307. package/skills/skills-codex/shared-references/experiment-integrity.md +49 -0
  308. package/skills/skills-codex/shared-references/external-cadence.md +334 -0
  309. package/skills/skills-codex/shared-references/fan-out-pattern.md +375 -0
  310. package/skills/skills-codex/shared-references/injection-hygiene.md +132 -0
  311. package/skills/skills-codex/shared-references/integration-contract.md +372 -0
  312. package/skills/skills-codex/shared-references/output-composition.md +98 -0
  313. package/skills/skills-codex/shared-references/output-language.md +45 -0
  314. package/skills/skills-codex/shared-references/output-manifest.md +40 -0
  315. package/skills/skills-codex/shared-references/output-versioning.md +111 -0
  316. package/skills/skills-codex/shared-references/patent-format-cn.md +199 -0
  317. package/skills/skills-codex/shared-references/patent-format-ep.md +173 -0
  318. package/skills/skills-codex/shared-references/patent-format-us.md +161 -0
  319. package/skills/skills-codex/shared-references/patent-writing-principles.md +197 -0
  320. package/skills/skills-codex/shared-references/prior-art-databases.md +141 -0
  321. package/skills/skills-codex/shared-references/resumable-runs.md +125 -0
  322. package/skills/skills-codex/shared-references/review-scope-limits.md +81 -0
  323. package/skills/skills-codex/shared-references/review-tracing.md +144 -0
  324. package/skills/skills-codex/shared-references/reviewer-independence.md +66 -0
  325. package/skills/skills-codex/shared-references/reviewer-routing.md +128 -0
  326. package/skills/skills-codex/shared-references/skill-governance.md +119 -0
  327. package/skills/skills-codex/shared-references/taste-calibration.md +90 -0
  328. package/skills/skills-codex/shared-references/venue-checklists.md +73 -0
  329. package/skills/skills-codex/shared-references/wiki-helper-resolution.md +69 -0
  330. package/skills/skills-codex/shared-references/writing-principles.md +525 -0
  331. package/skills/skills-codex/slides-polish/SKILL.md +563 -0
  332. package/skills/skills-codex/specification-writing/SKILL.md +211 -0
  333. package/skills/skills-codex/system-profile/SKILL.md +103 -0
  334. package/skills/skills-codex/training-check/SKILL.md +83 -0
  335. package/skills/skills-codex/vast-gpu/SKILL.md +394 -0
  336. package/skills/skills-codex/web-debug-search/SKILL.md +334 -0
  337. package/skills/skills-codex/wiki-enrich/SKILL.md +255 -0
  338. package/skills/skills-codex/writing-systems-papers/SKILL.md +184 -0
  339. package/skills/skills-codex-claude-review/README.md +79 -0
  340. package/skills/skills-codex-claude-review/README_CN.md +78 -0
  341. package/skills/skills-codex-claude-review/auto-paper-improvement-loop/SKILL.md +581 -0
  342. package/skills/skills-codex-claude-review/auto-review-loop/SKILL.md +510 -0
  343. package/skills/skills-codex-claude-review/novelty-check/SKILL.md +102 -0
  344. package/skills/skills-codex-claude-review/paper-figure/SKILL.md +319 -0
  345. package/skills/skills-codex-claude-review/paper-plan/SKILL.md +287 -0
  346. package/skills/skills-codex-claude-review/paper-write/SKILL.md +420 -0
  347. package/skills/skills-codex-claude-review/research-refine/SKILL.md +732 -0
  348. package/skills/skills-codex-claude-review/research-review/SKILL.md +149 -0
  349. package/skills/skills-codex-gemini-review/README.md +176 -0
  350. package/skills/skills-codex-gemini-review/README_CN.md +175 -0
  351. package/skills/skills-codex-gemini-review/auto-paper-improvement-loop/SKILL.md +331 -0
  352. package/skills/skills-codex-gemini-review/auto-review-loop/SKILL.md +304 -0
  353. package/skills/skills-codex-gemini-review/grant-proposal/SKILL.md +630 -0
  354. package/skills/skills-codex-gemini-review/idea-creator/SKILL.md +263 -0
  355. package/skills/skills-codex-gemini-review/idea-discovery/SKILL.md +275 -0
  356. package/skills/skills-codex-gemini-review/idea-discovery-robot/SKILL.md +365 -0
  357. package/skills/skills-codex-gemini-review/novelty-check/SKILL.md +92 -0
  358. package/skills/skills-codex-gemini-review/paper-figure/SKILL.md +289 -0
  359. package/skills/skills-codex-gemini-review/paper-plan/SKILL.md +265 -0
  360. package/skills/skills-codex-gemini-review/paper-poster-html/SKILL.md +102 -0
  361. package/skills/skills-codex-gemini-review/paper-slides/SKILL.md +582 -0
  362. package/skills/skills-codex-gemini-review/paper-write/SKILL.md +346 -0
  363. package/skills/skills-codex-gemini-review/paper-writing/SKILL.md +312 -0
  364. package/skills/skills-codex-gemini-review/research-refine/SKILL.md +674 -0
  365. package/skills/skills-codex-gemini-review/research-review/SKILL.md +112 -0
  366. package/skills/slides-polish/SKILL.md +565 -0
  367. package/skills/specification-writing/SKILL.md +211 -0
  368. package/skills/system-profile/SKILL.md +103 -0
  369. package/skills/training-check/SKILL.md +132 -0
  370. package/skills/vast-gpu/SKILL.md +394 -0
  371. package/skills/web-debug-search/SKILL.md +334 -0
  372. package/skills/wiki-enrich/SKILL.md +257 -0
  373. package/skills/writing-systems-papers/SKILL.md +184 -0
  374. package/templates/CLAUDE_MD_TEMPLATE.md +29 -0
  375. package/templates/EXPERIMENT_LOG_TEMPLATE.md +47 -0
  376. package/templates/EXPERIMENT_PLAN_TEMPLATE.md +51 -0
  377. package/templates/EXPERIMENT_PLAN_TEMPLATE_CN.md +53 -0
  378. package/templates/FINDINGS_TEMPLATE.md +52 -0
  379. package/templates/IDEA_CANDIDATES_TEMPLATE.md +47 -0
  380. package/templates/IDEA_CANDIDATES_TEMPLATE_CN.md +47 -0
  381. package/templates/INVENTION_BRIEF_TEMPLATE.md +87 -0
  382. package/templates/MANIFEST_TEMPLATE.md +7 -0
  383. package/templates/NARRATIVE_REPORT_TEMPLATE.md +49 -0
  384. package/templates/PAPER_PLAN_TEMPLATE.md +47 -0
  385. package/templates/PATENT_CLAIMS_TEMPLATE.md +78 -0
  386. package/templates/PATENT_SPECIFICATION_TEMPLATE.md +68 -0
  387. package/templates/README.md +57 -0
  388. package/templates/RESEARCH_BRIEF_TEMPLATE.md +35 -0
  389. package/templates/RESEARCH_BRIEF_TEMPLATE_CN.md +41 -0
  390. package/templates/RESEARCH_CONTRACT_TEMPLATE.md +60 -0
  391. package/templates/claude-hooks/corpus_write_guard.json +16 -0
  392. package/templates/claude-hooks/corpus_write_guard.py +85 -0
  393. package/templates/claude-hooks/meta_logging.json +74 -0
  394. package/templates/gitignore-trace.txt +3 -0
  395. package/tools/__pycache__/check_skills_inventory.cpython-314.pyc +0 -0
  396. package/tools/arxiv_fetch.py +311 -0
  397. package/tools/capture_filter.py +126 -0
  398. package/tools/check_skills_inventory.py +273 -0
  399. package/tools/convert_skills_to_llm_chat.py +282 -0
  400. package/tools/copilot_native_evidence.py +818 -0
  401. package/tools/deepxiv_fetch.py +213 -0
  402. package/tools/evidence_check.py +212 -0
  403. package/tools/exa_search.py +425 -0
  404. package/tools/experiment_queue/README.md +118 -0
  405. package/tools/experiment_queue/build_manifest.py +44 -0
  406. package/tools/experiment_queue/queue_manager.py +44 -0
  407. package/tools/extract_paper_style.py +560 -0
  408. package/tools/figure_renderer.py +69 -0
  409. package/tools/forensics_gate.py +669 -0
  410. package/tools/generate_codex_claude_review_overrides.py +299 -0
  411. package/tools/idea_discovery_gate.py +256 -0
  412. package/tools/install_aris.ps1 +1372 -0
  413. package/tools/install_aris.sh +1370 -0
  414. package/tools/install_aris_codex.sh +1023 -0
  415. package/tools/install_aris_copilot.sh +1052 -0
  416. package/tools/iteration_log.py +143 -0
  417. package/tools/lint_skills_helpers.sh +84 -0
  418. package/tools/meta_opt/check_ready.sh +80 -0
  419. package/tools/meta_opt/log_event.sh +91 -0
  420. package/tools/meta_opt/trigger_eval.py +280 -0
  421. package/tools/meta_opt/trigger_evals.sample.json +28 -0
  422. package/tools/openalex_fetch.py +326 -0
  423. package/tools/overleaf_audit.sh +104 -0
  424. package/tools/overleaf_setup.sh +150 -0
  425. package/tools/paper_illustration_image2.py +62 -0
  426. package/tools/provenance.py +294 -0
  427. package/tools/research_wiki.py +1720 -0
  428. package/tools/review_gate.py +502 -0
  429. package/tools/run_state.py +399 -0
  430. package/tools/save_trace.sh +477 -0
  431. package/tools/semantic_scholar_fetch.py +438 -0
  432. package/tools/skill-groups.tsv +116 -0
  433. package/tools/skill_picker.py +238 -0
  434. package/tools/smart_update.ps1 +521 -0
  435. package/tools/smart_update.sh +591 -0
  436. package/tools/smart_update_codex.sh +419 -0
  437. package/tools/smart_update_copilot.sh +605 -0
  438. package/tools/threat_scan.py +222 -0
  439. package/tools/verify_paper_audits.sh +487 -0
  440. package/tools/verify_papers.py +613 -0
  441. package/tools/verify_wiki_coverage.sh +176 -0
  442. package/tools/watchdog.py +485 -0
@@ -0,0 +1,437 @@
1
+ ---
2
+ name: kill-argument
3
+ description: "Two-thread adversarial review: a fresh reviewer constructs the strongest 200-word rejection memo, then a second fresh reviewer defends the paper point-by-point and surfaces still-unresolved critical issues. Use when user says \"kill argument\", \"adversarial review\", \"hostile review\", \"rebuttal preparation\", \"reviewer-2 simulation\", or before submitting a theory paper that has already passed standard review rounds."
4
+ argument-hint: "[paper-directory]"
5
+ allowed-tools: Bash(*), Read, Write, Edit, Grep, Glob, mcp__codex__codex
6
+ ---
7
+
8
+ # Kill Argument Exercise: Adversarial Attack-Defense Review
9
+
10
+ > 🔒 **Do not wrap this skill in `/loop`, `/schedule`, or `CronCreate`.** It is
11
+ > verdict-bearing — it produces an adversarial accept/reject verdict (attack →
12
+ > adjudication). Re-firing it on a wall-clock timer adds no new signal (the
13
+ > attack changes only when the *paper* changes). Schedule the *external wait
14
+ > that precedes it* — draft stable → then run this **once** before submission.
15
+ > See
16
+ > [`shared-references/external-cadence.md`](../shared-references/external-cadence.md).
17
+
18
+ Stress-test the headline claims of a paper against the strongest possible rejection argument: **$ARGUMENTS**
19
+
20
+ ## Why This Exists
21
+
22
+ Standard score-based reviews (`/research-review`, `/auto-paper-improvement-loop`) tend to produce **balanced** weakness lists. Each weakness gets ~equal attention, ranked CRITICAL > MAJOR > MINOR. Empirically, this misses one specific failure mode: the **single most damaging argument** a reviewer would write in a rejection paragraph — the one sentence that, if a senior area chair reads it, kills the paper.
23
+
24
+ A balanced reviewer might list "scope-overclaim risk" as MAJOR alongside 3-5 other MAJORs, never quite committing. An adversarial reviewer **must commit**: their entire job is to convince the area chair to reject in 200 words.
25
+
26
+ This skill runs that adversarial pass deliberately, then forces a second fresh reviewer to defend point-by-point, classify each rejection as already-fixed / partially-fixed / still-unresolved, and surface what's actually load-bearing.
27
+
28
+ **Empirical motivation:** in a real submission run, after several rounds of standard improvement (score 7-8/10), the kill-argument exercise surfaced framing weaknesses that no prior review caught (e.g., a setting being mostly conditional rather than truly general, or a baseline being irrelevant to real systems). Author rebuttal forced explicit scope qualifications in abstract and discussion that weren't visible from the score-based reviews alone.
29
+
30
+ ## How This Differs From Other Review Skills
31
+
32
+ | Skill | What it asks the reviewer | Output |
33
+ |-------|---------------------------|--------|
34
+ | Standard peer review | "Score this paper, list weaknesses by severity" | balanced weakness list |
35
+ | `/research-review` | "Deep technical review of methods + claims" | structured deep critique |
36
+ | `/proof-checker` | "Is this theorem actually proved?" | per-step proof obligation audit |
37
+ | `/paper-claim-audit` | "Does the paper report numbers truthfully?" | per-claim evidence verification |
38
+ | `/citation-audit` | "Are citations real and used in correct context?" | per-entry KEEP/FIX/REPLACE/REMOVE |
39
+ | **`/kill-argument`** | **"Write the single strongest rejection paragraph; then defend it."** | **attack memo + per-point defense + unresolved surfaced** |
40
+
41
+ This skill is **complementary**, not a replacement. Run after standard reviews when you want to know what the worst-case reviewer paragraph would look like, before camera-ready or rebuttal preparation.
42
+
43
+ ## When To Use
44
+
45
+ - After 1-2 rounds of `/auto-paper-improvement-loop` settled at a stable score, but before submission. Surfaces what additional fixes would close the headline-attack gap.
46
+ - During rebuttal preparation, to predict reviewer-2's strongest objection so you can prepare the response in advance.
47
+ - For theory papers with a high-level title that may oversimplify the actual theorem (the most common reject-attack pattern).
48
+ - For papers where a reviewer might attack scope, assumption-vs-claim mismatch, missing proof obligations, or evidence-vs-headline gaps.
49
+
50
+ This skill is most valuable for **theory papers** with ≥5 theorem-class environments (so the headline depends on real proof obligations). For empirical papers without theorems, use `/research-review` instead.
51
+
52
+ ## Constants
53
+
54
+ - **REVIEWER_MODEL** = `gpt-5.6-sol` (default; `gpt-5.5` is the capability fallback, `gpt-5.4` only as an explicit legacy override). Reviewer reasoning effort = `ultra` for the attack / defense / adjudication threads (deep-audit tier; capability fallback per `shared-references/reviewer-routing.md`, never below `xhigh`). Beast-mode axis probes stay at `xhigh`.
55
+ - **CONTEXT_POLICY** = `fresh` (REVIEWER_BIAS_GUARD). Each thread is a fresh `mcp__codex__codex` call. **Never** use `mcp__codex__codex-reply`. No prior review summary, fix list, or executor explanation enters either prompt.
56
+ - **ATTACK_LENGTH** = approximately 200 words (do not exceed 250). Single coherent argument, not a list.
57
+ - **DEFENSE_DECOMPOSITION** = 3-7 atomic rejection points extracted from the attack memo. Each gets its own classification.
58
+ - **CLASSIFICATION** = `answered_by_current_text` / `partially_answered` / `still_unresolved`. (Names chosen so the adjudicator does not assume "fixed" implies prior history of patching — they read the paper as a fresh reviewer would.)
59
+ - **OUTPUT** = `KILL_ARGUMENT.md` (human-readable) + `KILL_ARGUMENT.json` (machine-readable) in the paper directory.
60
+ - **RENDER_HTML = true** — When `true` (default), auto-render `KILL_ARGUMENT.md` to HTML after writing the report. Uses **full Codex review gate** (audit-class artifact — full render-fidelity check matches the skill's cross-model audit invariant; the sidecar `KILL_ARGUMENT.json` is also passed to the renderer). Set `false` to skip, or pass `— render html: false`.
61
+
62
+ ## Workflow
63
+
64
+ ### Step 1: Discover paper files
65
+
66
+ Locate the paper directory and inventory the source.
67
+
68
+ ```bash
69
+ PAPER_DIR="$ARGUMENTS" # e.g., paper-overleaf/ or paper/
70
+ cd "$PAPER_DIR"
71
+
72
+ # Find the LaTeX entry point
73
+ ENTRY=$(grep -lE '^\\documentclass' *.tex 2>/dev/null | head -1)
74
+ echo "Entry: $ENTRY"
75
+
76
+ # Find all source files codex should read
77
+ find . -name "*.tex" -not -path "./.git/*" 2>/dev/null
78
+ find . -name "*.bib" -not -path "./.git/*" 2>/dev/null
79
+ find figures/ -name "*.pdf" -o -name "*.png" 2>/dev/null
80
+ ls -la *.pdf 2>/dev/null # compiled PDF
81
+ ```
82
+
83
+ If a compiled PDF is missing, the skill should still run on .tex source alone, but the prompt should mention this so the reviewer doesn't waste cycles trying to extract from a non-existent PDF.
84
+
85
+ ### Step 2: Attack memo (Thread 1, fresh codex)
86
+
87
+ Invoke `mcp__codex__codex` (NOT `codex-reply`) with the following prompt structure:
88
+
89
+ ```
90
+ mcp__codex__codex:
91
+ model: gpt-5.6-sol
92
+ config: {"model_reasoning_effort": "ultra"}
93
+ sandbox: read-only
94
+ cwd: <paper directory>
95
+ prompt: |
96
+ You are simulating a hostile NeurIPS / ICLR / ICML reviewer for a paper.
97
+ This is a kill-argument adversarial check — your task is NOT to give a
98
+ balanced review but to construct the **single strongest argument for
99
+ rejecting this paper**.
100
+
101
+ ## Files to read
102
+ - LaTeX entry: <ENTRY>
103
+ - All section files under sections/ or wherever they live
104
+ - Macro files (math_commands.tex, etc.)
105
+ - Compiled PDF: <main.pdf> (if available)
106
+
107
+ Read the source carefully. Do not consult any prior reviews, fix lists,
108
+ or summaries; this must be a fresh, zero-context adversarial pass.
109
+
110
+ ## Your task
111
+ Construct the single best argument to reject this paper in approximately
112
+ 200 words. Your goal is to write the worst-case rejection memo a senior
113
+ NeurIPS area chair would produce after reading the paper.
114
+
115
+ Focus on these axes (pick the most damaging combination, do not list all):
116
+ 1. Theorem validity: are central theorems actually proved as stated?
117
+ 2. Assumption-vs-claim mismatch: does the body silently retreat to a
118
+ narrower object than the title/abstract advertise?
119
+ 3. Missing proof obligations: is a fundamental lemma invoked but not
120
+ proved (e.g., concentration, generic position, prefactor envelope)
121
+ that the headline depends on?
122
+ 4. Limit-order ambiguity: are limits in K/n/d/eps composed in a way the
123
+ paper does not commit to?
124
+ 5. Claim-vs-evidence gap: is the empirical/numerical evidence too narrow
125
+ to support the breadth of the stated theorem or take-away?
126
+ 6. Scope overclaim: does the title or abstract sell a result substantially
127
+ broader than what the body proves?
128
+
129
+ ## Constraints
130
+ - Approximately 200 words total (do NOT exceed 250).
131
+ - Single argument, not a list — pick the most damaging line of attack
132
+ and develop it.
133
+ - Cite specific file:line locations or equation numbers when accusing.
134
+ - Tone: dispassionate but uncompromising. Do NOT hedge. Do NOT acknowledge
135
+ mitigations the paper might have made elsewhere. This is the rejection
136
+ paragraph; the defense gets the next pass.
137
+ - Do NOT reference prior review rounds, fix lists, or any context outside
138
+ the current paper files.
139
+
140
+ Output: just the rejection memo, nothing else.
141
+ ```
142
+
143
+ Save the returned `threadId` for the trace; do NOT pass it to Thread 2. Save the attack memo verbatim — both Thread 2 and the human-readable report use it.
144
+
145
+ ### Step 2.5 (optional, `beast` effort): multi-axis attack fan-out
146
+
147
+ **Default OFF.** The deliverable of this skill *is* a verdict — the single
148
+ strongest rejection paragraph — and
149
+ [`shared-references/fan-out-pattern.md`](../shared-references/fan-out-pattern.md)
150
+ is explicit: **do not fan out the verdict; fan out only the evidence that
151
+ feeds it.** The default single-commitment attack (Step 2) is deliberate —
152
+ forcing one paragraph produces sharper feedback than a balanced list (see *Why
153
+ This Exists*). Do **not** replace it with a list.
154
+
155
+ Under `beast` effort you may widen the *evidence* the commitment draws on
156
+ without diluting the commitment:
157
+
158
+ 1. **Axis probes (evidence breadth).** Run the six attack axes (theorem
159
+ validity / assumption-vs-claim / missing obligation / limit-order /
160
+ claim-vs-evidence / scope-overclaim) as **separate fresh-codex probes**,
161
+ each asked for the strongest ~120-word thrust *on that axis alone*. These
162
+ are evidence-gathering, not the verdict. Probes run at `xhigh` (not
163
+ `ultra`) — six serial delegating calls would multiply cost for evidence
164
+ that the ultra-tier commit re-judges anyway.
165
+ - **These are NOT Claude subagents, and there is deliberately NO `Agent`
166
+ grant.** Each probe is a fresh `mcp__codex__codex` call — the adversary
167
+ must be cross-model (non-Claude). Codex MCP is **serial** (concurrent
168
+ codex calls hang), so the probes run **sequentially** — Tier-3 in the
169
+ fan-out ladder. This is exactly why `kill-argument` lists no `Agent` in
170
+ `allowed-tools`: it spawns nothing; it threads codex calls.
171
+ 2. **Commit (the verdict, still single).** A final fresh-codex synthesis reads
172
+ the six probes plus the paper and must **commit to the single most damaging
173
+ ~200-word rejection paragraph** — selecting and fusing at most two axes, NOT
174
+ listing all six. The Step-2 commitment requirement is unchanged; the probes
175
+ only ensure no axis was overlooked before committing.
176
+
177
+ The adjudication (Step 3) then runs against this committed attack exactly as in
178
+ the default flow. Cost: `beast` adds ~6 extra serial codex calls — use it for
179
+ the final pre-submission pass on a high-stakes paper, not routinely.
180
+
181
+ Tracing: record each probe's `threadId` (`axis_probe_thread_ids[]`) and the
182
+ synthesis `threadId` in the trace, the same way Steps 2–3 save their thread
183
+ ids. The committed attack memo, not the six probes, is what Step 3 consumes.
184
+
185
+ ### Step 3: Adjudication memo (Thread 2, fresh codex with attack + paper)
186
+
187
+ Invoke a second `mcp__codex__codex` call (still NOT `codex-reply` — Thread 2 is independent of Thread 1's codex history):
188
+
189
+ ```
190
+ mcp__codex__codex:
191
+ model: gpt-5.6-sol
192
+ config: {"model_reasoning_effort": "ultra"}
193
+ sandbox: read-only
194
+ cwd: <paper directory>
195
+ prompt: |
196
+ You are an independent area-chair adjudicator examining whether the
197
+ current paper text answers a hostile reviewer's rejection memo.
198
+ You are NOT the paper's defender — your job is to read the attack
199
+ point-by-point and rule, from the current source files alone,
200
+ whether each point stands or falls. Fresh, zero-context adjudication;
201
+ do not reference any prior reviews / fix lists.
202
+
203
+ ## Paper files
204
+ [list paths same as Step 2]
205
+
206
+ ## The hostile reviewer's rejection memo (the "attack")
207
+ > <attack memo verbatim from Thread 1>
208
+
209
+ ## Your task
210
+ The attack is one continuous argument, but it makes multiple distinct
211
+ rejection points that you must adjudicate separately. Decompose the
212
+ attack into its atomic rejection points (3-7 of them), then for each
213
+ point classify it:
214
+
215
+ - answered_by_current_text: the current paper source already mitigates
216
+ this point (cite specific file:line evidence)
217
+ - partially_answered: paper has some response but not enough to refute
218
+ the attack as written
219
+ - still_unresolved: paper has no effective response
220
+
221
+ The label `answered_by_current_text` is intentional — "fixed" implies
222
+ history of patching and biases toward optimism. You are reading the
223
+ paper as a reviewer would, with no knowledge of prior round drafts.
224
+
225
+ For each rejection point, output:
226
+ ### Point P_n: <short label>
227
+ **Attack claim**: <the specific accusation, ~30 words>
228
+ **Verdict**: answered_by_current_text | partially_answered | still_unresolved
229
+ **Evidence (or lack of)**: <cite file:line, ~50 words>
230
+ **Severity if unresolved**: critical | major | minor
231
+ **If unresolved, recommended fix**: <one specific actionable sentence>
232
+
233
+ After per-point analysis, output:
234
+
235
+ ## Summary
236
+ Total rejection points: N
237
+ - answered_by_current_text: X
238
+ - partially_answered: Y
239
+ - still_unresolved: Z
240
+
241
+ ## Net assessment
242
+ <one short paragraph: would this paper survive a senior area-chair read
243
+ of the attack memo, given only what is in the current source? Be honest —
244
+ if Y or Z > 0 and they hit the headline, say so.>
245
+
246
+ ## Top action items (in priority order, max 3)
247
+ 1. ...
248
+ 2. ...
249
+ 3. ...
250
+
251
+ ## Constraints
252
+ - Do NOT consult any prior round reviews or fix lists. Adjudication must
253
+ be made strictly from current paper files.
254
+ - If the paper cannot refute a point, do NOT minimize — keep severity
255
+ honest.
256
+ - If a point reflects an author-chosen position (e.g., conscious title
257
+ scope decision), classify as `partially_answered` with a note that the
258
+ position is intentional, AND say whether this position is sustainable
259
+ under the attack — do NOT auto-grade as `answered_by_current_text`
260
+ just because it is intentional.
261
+ - Be specific. No flattery, no hedging, no rationalizing on the paper's
262
+ behalf.
263
+ ```
264
+
265
+ Save the returned `threadId`.
266
+
267
+ ### Step 4: Write KILL_ARGUMENT.md and KILL_ARGUMENT.json
268
+
269
+ Compose the human-readable report `<paper-dir>/KILL_ARGUMENT.md`:
270
+
271
+ ```markdown
272
+ # Kill Argument Report — <paper title>
273
+
274
+ **Date**: <YYYY-MM-DD>
275
+ **Reviewer model**: <resolved pair that actually ran — target gpt-5.6-sol ultra>, fresh threads (no codex-reply)
276
+ **Attack thread**: <threadId 1>
277
+ **Adjudicator thread**: <threadId 2>
278
+ **Verdict**: <PASS / WARN / FAIL / NOT_APPLICABLE / BLOCKED / ERROR> (`reason_code: <...>`)
279
+
280
+ ## Net assessment
281
+
282
+ <paragraph from adjudicator memo's "Net assessment">
283
+
284
+ ## Attack memo (verbatim)
285
+
286
+ > <attack memo from Thread 1>
287
+
288
+ ## Adjudication (per-point)
289
+
290
+ <copy verbatim from Thread 2 — uses labels answered_by_current_text / partially_answered / still_unresolved>
291
+
292
+ ## Top action items
293
+
294
+ <copy from Thread 2>
295
+
296
+ ## Recommendation
297
+
298
+ If P_4 (or whatever still_unresolved critical) is research-level, record
299
+ it as a known open problem in the conclusion / limitations. If it is
300
+ writing-level, queue for next /auto-paper-improvement-loop round.
301
+ ```
302
+
303
+ Compose the machine-readable `<paper-dir>/KILL_ARGUMENT.json` per the
304
+ ARIS Audit Artifact Schema (`shared-references/assurance-contract.md`):
305
+
306
+ ```json
307
+ {
308
+ "audit_skill": "kill-argument",
309
+ "verdict": "PASS | WARN | FAIL | NOT_APPLICABLE | BLOCKED | ERROR",
310
+ "reason_code": "<see verdict mapping below>",
311
+ "summary": "<one-line summary, ~80 chars>",
312
+ "audited_input_hashes": {
313
+ "main.tex": "sha256:<...>",
314
+ "sec/0.abstract.tex": "sha256:<...>",
315
+ "sec/<each-section>.tex": "sha256:<...>",
316
+ "references.bib": "sha256:<...>",
317
+ "main.pdf": "sha256:<...>"
318
+ },
319
+ "trace_path": ".aris/traces/kill-argument/<date>_run<NN>/",
320
+ "thread_id": "<defense threadId — primary; attack threadId in details>",
321
+ "reviewer_model": "<resolved — the model that actually ran (target: gpt-5.6-sol)>",
322
+ "reviewer_reasoning": "<resolved — the effort that actually ran (target: ultra)>",
323
+ "generated_at": "<UTC ISO-8601>",
324
+ "details": {
325
+ "attack_thread_id": "<threadId 1>",
326
+ "defense_thread_id": "<threadId 2 — same as top-level thread_id>",
327
+ "attack_memo": "<verbatim>",
328
+ "decomposed_points": [
329
+ {
330
+ "id": "P_1",
331
+ "label": "<short label>",
332
+ "attack_claim": "<...>",
333
+ "verdict": "answered_by_current_text | partially_answered | still_unresolved",
334
+ "evidence": "<file:line citation>",
335
+ "severity_if_unresolved": "critical | major | minor",
336
+ "recommended_fix": "<...>"
337
+ }
338
+ ],
339
+ "counts": {
340
+ "answered_by_current_text": <int>,
341
+ "partially_answered": <int>,
342
+ "still_unresolved": <int>
343
+ },
344
+ "net_assessment": "<adjudicator memo's net assessment>",
345
+ "top_action_items": ["...", "...", "..."]
346
+ }
347
+ }
348
+ ```
349
+
350
+ **Hash inputs** (`audited_input_hashes`): use paper-relative paths,
351
+ `sha256` of every `.tex` consumed plus `references.bib` and the
352
+ compiled `main.pdf` if it exists. The verifier rehashes these on
353
+ `verify_paper_audits.sh` and flags `STALE` if the user edited the
354
+ paper after running the audit.
355
+
356
+ **Verdict mapping** (every (counts, severity) tuple must hit exactly one row):
357
+
358
+ | Verdict | reason_code | Trigger |
359
+ |---|---|---|
360
+ | `FAIL` | `unresolved_critical` | ≥1 `still_unresolved` at `critical` severity |
361
+ | `WARN` | `unresolved_major_or_minor` | ≥1 `still_unresolved` at `major` or `minor` severity (and no `critical`) |
362
+ | `WARN` | `partial_critical_or_repeated_major` | 0 `still_unresolved`, AND ≥1 `partially_answered` at `critical` or `major` |
363
+ | `PASS` | `defense_survives_with_minor_partial_only` | 0 `still_unresolved`, AND ≥1 `partially_answered`, all at `minor` severity |
364
+ | `PASS` | `defense_survives` | 0 `still_unresolved`, AND 0 `partially_answered` |
365
+ | `NOT_APPLICABLE` | `not_theory_or_scope_paper` | Paper has <2 `\begin{theorem\|lemma\|proposition\|corollary}` AND no scope / generality claims in abstract |
366
+ | `NOT_APPLICABLE` | `headline_unstable` | Title or abstract changed within the last 2 commits — re-run after headline stabilizes |
367
+ | `BLOCKED` | `paper_compile_failed` | Compiled PDF missing AND `main.tex` does not compile clean — adjudication needs source fidelity |
368
+ | `BLOCKED` | `source_files_missing` | `main.tex` not found, or no `sec/*.tex` files |
369
+ | `ERROR` | `codex_api_error` | `mcp__codex__codex` call failed |
370
+ | `ERROR` | `decomposition_parse_failed` | Adjudicator thread did not return parseable per-point structure |
371
+ | `ERROR` | `trace_save_failed` | Trace directory write failed |
372
+
373
+ `PASS` requires `still_unresolved == 0`. With `still_unresolved == 0`, any
374
+ `partially_answered` at `major` or higher makes the best available verdict
375
+ `WARN` — never `PASS`.
376
+
377
+ The verdict is computed from the per-point counts; do NOT let the
378
+ defense thread output the top-level verdict directly (that would let
379
+ it self-grade). The skill code does the verdict mapping.
380
+
381
+ ### Step 5: Print summary
382
+
383
+ To the user:
384
+
385
+ ```
386
+ 🗡 Kill Argument complete.
387
+
388
+ Attack: <one-sentence summary of the rejection thrust>
389
+
390
+ Adjudication breakdown:
391
+ answered_by_current_text: X
392
+ partially_answered: Y
393
+ still_unresolved: Z ← critical: <names>
394
+
395
+ Verdict: <PASS / WARN / FAIL / NOT_APPLICABLE / BLOCKED / ERROR>
396
+ Reason: <reason_code, e.g., defense_survives, unresolved_critical>
397
+
398
+ Top action items:
399
+ 1. ...
400
+ 2. ...
401
+ 3. ...
402
+
403
+ Full report: <paper-dir>/KILL_ARGUMENT.md
404
+ ```
405
+
406
+ ## Output Contract
407
+
408
+ - `<paper-dir>/KILL_ARGUMENT.md` — human-readable report
409
+ - `<paper-dir>/KILL_ARGUMENT.json` — machine-readable ledger
410
+ - `.aris/traces/kill-argument/<date>_runNN/` — per-thread codex traces (Attack memo + Adjudication memo)
411
+ - Optional: applied fixes if user explicitly requests; default is **detect-only, do not auto-modify**.
412
+ - `<paper-dir>/KILL_ARGUMENT.html` (when `RENDER_HTML = true`, default) — single-file HTML view auto-rendered via `/render-html "<paper-dir>/KILL_ARGUMENT.md" --json "<paper-dir>/KILL_ARGUMENT.json"`. Full review gate applies. The `.review.json` sidecar carries the render-fidelity verdict. **Non-blocking**: if `/render-html` fails (helper missing, Codex MCP unavailable, file write error), log the failure and treat the skill as complete — the HTML view is a convenience, not a prerequisite for the kill-argument verdict.
413
+
414
+ ## Key Rules
415
+
416
+ - **Fresh thread per call.** Both Attack and Adjudication use `mcp__codex__codex`, never `codex-reply`. Thread 1 and Thread 2 must not share codex context.
417
+ - **Zero prior context.** Neither thread receives prior round reviews, fix lists, executor summaries, or improvement-loop logs.
418
+ - **Attack must commit.** Single argument, ~200 words. No "consider also" hedge. The whole value is in forcing the reviewer to pick the most damaging line.
419
+ - **Adjudicator must classify, not minimize.** `still_unresolved` is honest if the paper has no effective response. Don't downgrade to `partially_answered` unless evidence is real.
420
+ - **Author-chosen positions** (e.g., deliberate title scope, deliberate omission of qualifier): mark `partially_answered` with note that the position is intentional, AND say whether the position is sustainable under the attack. Don't auto-grade as `answered_by_current_text` just because it's intentional.
421
+ - **Verdict is computed by the skill, not by the adjudicator.** The Codex thread emits per-point classifications; the skill code maps those to one of the 6 audit verdicts via the table in Step 4. Never let the adjudicator self-grade the top-level verdict.
422
+ - **Detect-only by direct invocation; can be invoked by `/auto-paper-improvement-loop` Step 5.5 which then merges unresolved findings into its fix list.** When a user runs `/kill-argument paper/` directly, the output is informational and the human decides whether to act. When the skill is invoked from inside the auto-improvement loop, the loop reads `KILL_ARGUMENT.json`, deduplicates against its existing weakness list, and feeds novel `still_unresolved` points into Step 6 fixes — `/kill-argument` itself never edits paper files.
423
+
424
+ ## When NOT to Use
425
+
426
+ - Empirical papers without theorems / scope claims — `/research-review` is more useful. The skill emits `NOT_APPLICABLE` with `reason_code: not_theory_or_scope_paper` in this case.
427
+ - Very early drafts where the headline isn't stable yet — fix the headline first. The skill emits `NOT_APPLICABLE` with `reason_code: headline_unstable` if the title or abstract changed within the last 2 commits.
428
+ - Papers with ongoing experiments — wait until results stabilize, then run.
429
+ - (`/auto-paper-improvement-loop` Step 5.5 used to run this protocol inline; as of May 2026 it now invokes `/kill-argument` and reads `KILL_ARGUMENT.json` instead, so there is no longer a "do not invoke from inside auto-loop" exclusion.)
430
+
431
+ ## Review Tracing
432
+
433
+ After each `mcp__codex__codex` reviewer call, save the trace following `shared-references/review-tracing.md` (Policy C — forensic; never silently skip). Use `save_trace.sh` (resolved per the chain in `shared-references/integration-contract.md` §2) or write files directly to `.aris/traces/kill-argument/<date>_run<NN>/`. Both threads' raw responses should be preserved.
434
+
435
+ ## Notes
436
+
437
+ This skill was extracted as a standalone primitive from `/auto-paper-improvement-loop` Step 5.5 in May 2026, after the protocol proved valuable in surfacing headline-vs-body scope gaps that score-based reviews missed. The attack-then-defense pattern was kept exactly because of empirical evidence that asking one model to "write the rejection memo" produces qualitatively different feedback than asking it to "review and grade" — the former forces commitment, the latter encourages hedging.