@humanbased/crosscheck 1.2.0 → 1.3.0-beta.100

This diff represents the content of publicly available package versions that have been released to one of the supported registries. The information contained in this diff is provided for informational purposes only and reflects changes between package versions as they appear in their respective public registries.
Files changed (438) hide show
  1. package/LICENSE +1 -1
  2. package/README.md +183 -375
  3. package/README.zh.md +1 -1
  4. package/assets/icon-256.png +0 -0
  5. package/assets/linear-comment.svg +18 -0
  6. package/assets/linear-onboard.svg +30 -0
  7. package/assets/linear-status.svg +23 -0
  8. package/assets/linear-test.svg +34 -0
  9. package/assets/skills/code-review/.crosscheck-skill.json +9 -0
  10. package/assets/skills/code-review/LICENSE +21 -0
  11. package/assets/skills/code-review/SKILL.md +89 -0
  12. package/assets/skills/code-review/agents/openai.yaml +3 -0
  13. package/assets/skills/code-review-skill/.crosscheck-skill.json +9 -0
  14. package/assets/skills/code-review-skill/LICENSE +21 -0
  15. package/assets/skills/code-review-skill/SKILL.md +231 -0
  16. package/assets/skills/code-review-skill/assets/pr-review-template.md +137 -0
  17. package/assets/skills/code-review-skill/assets/review-checklist.md +123 -0
  18. package/assets/skills/code-review-skill/reference/angular.md +768 -0
  19. package/assets/skills/code-review-skill/reference/architecture-review-guide.md +472 -0
  20. package/assets/skills/code-review-skill/reference/c.md +890 -0
  21. package/assets/skills/code-review-skill/reference/code-quality-universal.md +488 -0
  22. package/assets/skills/code-review-skill/reference/code-review-best-practices.md +136 -0
  23. package/assets/skills/code-review-skill/reference/common-bugs-checklist.md +286 -0
  24. package/assets/skills/code-review-skill/reference/cpp.md +893 -0
  25. package/assets/skills/code-review-skill/reference/cross-cutting/async-concurrency-patterns.md +515 -0
  26. package/assets/skills/code-review-skill/reference/cross-cutting/error-handling-principles.md +492 -0
  27. package/assets/skills/code-review-skill/reference/cross-cutting/n-plus-one-queries.md +309 -0
  28. package/assets/skills/code-review-skill/reference/cross-cutting/sql-injection-prevention.md +308 -0
  29. package/assets/skills/code-review-skill/reference/cross-cutting/xss-prevention.md +264 -0
  30. package/assets/skills/code-review-skill/reference/csharp.md +525 -0
  31. package/assets/skills/code-review-skill/reference/css-less-sass.md +661 -0
  32. package/assets/skills/code-review-skill/reference/django.md +985 -0
  33. package/assets/skills/code-review-skill/reference/fastapi.md +580 -0
  34. package/assets/skills/code-review-skill/reference/go.md +993 -0
  35. package/assets/skills/code-review-skill/reference/java.md +409 -0
  36. package/assets/skills/code-review-skill/reference/java8.md +586 -0
  37. package/assets/skills/code-review-skill/reference/kotlin.md +1018 -0
  38. package/assets/skills/code-review-skill/reference/nestjs.md +593 -0
  39. package/assets/skills/code-review-skill/reference/performance-review-guide.md +816 -0
  40. package/assets/skills/code-review-skill/reference/php.md +684 -0
  41. package/assets/skills/code-review-skill/reference/python.md +1073 -0
  42. package/assets/skills/code-review-skill/reference/qt.md +757 -0
  43. package/assets/skills/code-review-skill/reference/react.md +871 -0
  44. package/assets/skills/code-review-skill/reference/ruby.md +964 -0
  45. package/assets/skills/code-review-skill/reference/rust.md +846 -0
  46. package/assets/skills/code-review-skill/reference/security-review-guide.md +494 -0
  47. package/assets/skills/code-review-skill/reference/svelte.md +1064 -0
  48. package/assets/skills/code-review-skill/reference/swift.md +936 -0
  49. package/assets/skills/code-review-skill/reference/typescript.md +1016 -0
  50. package/assets/skills/code-review-skill/reference/vue.md +924 -0
  51. package/assets/skills/code-review-skill/reference/zig.md +440 -0
  52. package/assets/skills/code-review-skill/scripts/pr-analyzer.py +435 -0
  53. package/assets/skills/code-review-skill/scripts/test_pr_analyzer.py +380 -0
  54. package/assets/skills/codebase-design/.crosscheck-skill.json +9 -0
  55. package/assets/skills/codebase-design/DEEPENING.md +37 -0
  56. package/assets/skills/codebase-design/DESIGN-IT-TWICE.md +44 -0
  57. package/assets/skills/codebase-design/LICENSE +21 -0
  58. package/assets/skills/codebase-design/SKILL.md +114 -0
  59. package/assets/skills/codebase-design/agents/openai.yaml +3 -0
  60. package/assets/skills/diagnosing-bugs/.crosscheck-skill.json +9 -0
  61. package/assets/skills/diagnosing-bugs/LICENSE +21 -0
  62. package/assets/skills/diagnosing-bugs/SKILL.md +134 -0
  63. package/assets/skills/diagnosing-bugs/agents/openai.yaml +3 -0
  64. package/assets/skills/diagnosing-bugs/scripts/hitl-loop.template.sh +41 -0
  65. package/crosscheck.config.example.yml +124 -11
  66. package/dist/__tests__/adoption.test.d.ts +2 -0
  67. package/dist/__tests__/adoption.test.d.ts.map +1 -0
  68. package/dist/__tests__/adoption.test.js +175 -0
  69. package/dist/__tests__/adoption.test.js.map +1 -0
  70. package/dist/__tests__/auto-fix-branch.test.d.ts +2 -0
  71. package/dist/__tests__/auto-fix-branch.test.d.ts.map +1 -0
  72. package/dist/__tests__/auto-fix-branch.test.js +143 -0
  73. package/dist/__tests__/auto-fix-branch.test.js.map +1 -0
  74. package/dist/__tests__/board.test.js +13 -0
  75. package/dist/__tests__/board.test.js.map +1 -1
  76. package/dist/__tests__/can-write-verdict.test.d.ts +2 -0
  77. package/dist/__tests__/can-write-verdict.test.d.ts.map +1 -0
  78. package/dist/__tests__/can-write-verdict.test.js +31 -0
  79. package/dist/__tests__/can-write-verdict.test.js.map +1 -0
  80. package/dist/__tests__/codex-env.test.d.ts +2 -0
  81. package/dist/__tests__/codex-env.test.d.ts.map +1 -0
  82. package/dist/__tests__/codex-env.test.js +70 -0
  83. package/dist/__tests__/codex-env.test.js.map +1 -0
  84. package/dist/__tests__/codex.test.js +100 -25
  85. package/dist/__tests__/codex.test.js.map +1 -1
  86. package/dist/__tests__/comment-bodies.test.js +49 -1
  87. package/dist/__tests__/comment-bodies.test.js.map +1 -1
  88. package/dist/__tests__/commit-subject.test.d.ts +2 -0
  89. package/dist/__tests__/commit-subject.test.d.ts.map +1 -0
  90. package/dist/__tests__/commit-subject.test.js +72 -0
  91. package/dist/__tests__/commit-subject.test.js.map +1 -0
  92. package/dist/__tests__/conflict-resolve.test.js +44 -1
  93. package/dist/__tests__/conflict-resolve.test.js.map +1 -1
  94. package/dist/__tests__/credential-free-origin.test.d.ts +2 -0
  95. package/dist/__tests__/credential-free-origin.test.d.ts.map +1 -0
  96. package/dist/__tests__/credential-free-origin.test.js +59 -0
  97. package/dist/__tests__/credential-free-origin.test.js.map +1 -0
  98. package/dist/__tests__/detector.test.js +87 -1
  99. package/dist/__tests__/detector.test.js.map +1 -1
  100. package/dist/__tests__/error-classification.test.js +34 -0
  101. package/dist/__tests__/error-classification.test.js.map +1 -1
  102. package/dist/__tests__/fix.test.js +69 -0
  103. package/dist/__tests__/fix.test.js.map +1 -1
  104. package/dist/__tests__/harness-section.test.d.ts +2 -0
  105. package/dist/__tests__/harness-section.test.d.ts.map +1 -0
  106. package/dist/__tests__/harness-section.test.js +46 -0
  107. package/dist/__tests__/harness-section.test.js.map +1 -0
  108. package/dist/__tests__/kickass.test.js +31 -0
  109. package/dist/__tests__/kickass.test.js.map +1 -1
  110. package/dist/__tests__/linear-branding.test.d.ts +2 -0
  111. package/dist/__tests__/linear-branding.test.d.ts.map +1 -0
  112. package/dist/__tests__/linear-branding.test.js +156 -0
  113. package/dist/__tests__/linear-branding.test.js.map +1 -0
  114. package/dist/__tests__/linear-client.test.d.ts +2 -0
  115. package/dist/__tests__/linear-client.test.d.ts.map +1 -0
  116. package/dist/__tests__/linear-client.test.js +120 -0
  117. package/dist/__tests__/linear-client.test.js.map +1 -0
  118. package/dist/__tests__/linear-comment.test.d.ts +2 -0
  119. package/dist/__tests__/linear-comment.test.d.ts.map +1 -0
  120. package/dist/__tests__/linear-comment.test.js +151 -0
  121. package/dist/__tests__/linear-comment.test.js.map +1 -0
  122. package/dist/__tests__/linear-identity.test.d.ts +2 -0
  123. package/dist/__tests__/linear-identity.test.d.ts.map +1 -0
  124. package/dist/__tests__/linear-identity.test.js +253 -0
  125. package/dist/__tests__/linear-identity.test.js.map +1 -0
  126. package/dist/__tests__/linear-notify.test.d.ts +2 -0
  127. package/dist/__tests__/linear-notify.test.d.ts.map +1 -0
  128. package/dist/__tests__/linear-notify.test.js +144 -0
  129. package/dist/__tests__/linear-notify.test.js.map +1 -0
  130. package/dist/__tests__/linear-ref.test.d.ts +2 -0
  131. package/dist/__tests__/linear-ref.test.d.ts.map +1 -0
  132. package/dist/__tests__/linear-ref.test.js +261 -0
  133. package/dist/__tests__/linear-ref.test.js.map +1 -0
  134. package/dist/__tests__/linear-test-ref.test.d.ts +2 -0
  135. package/dist/__tests__/linear-test-ref.test.d.ts.map +1 -0
  136. package/dist/__tests__/linear-test-ref.test.js +81 -0
  137. package/dist/__tests__/linear-test-ref.test.js.map +1 -0
  138. package/dist/__tests__/linear-verify.test.d.ts +2 -0
  139. package/dist/__tests__/linear-verify.test.d.ts.map +1 -0
  140. package/dist/__tests__/linear-verify.test.js +132 -0
  141. package/dist/__tests__/linear-verify.test.js.map +1 -0
  142. package/dist/__tests__/linear-worker.test.d.ts +2 -0
  143. package/dist/__tests__/linear-worker.test.d.ts.map +1 -0
  144. package/dist/__tests__/linear-worker.test.js +83 -0
  145. package/dist/__tests__/linear-worker.test.js.map +1 -0
  146. package/dist/__tests__/linear-write-possible.test.d.ts +2 -0
  147. package/dist/__tests__/linear-write-possible.test.d.ts.map +1 -0
  148. package/dist/__tests__/linear-write-possible.test.js +30 -0
  149. package/dist/__tests__/linear-write-possible.test.js.map +1 -0
  150. package/dist/__tests__/loader.test.js +2 -2
  151. package/dist/__tests__/loader.test.js.map +1 -1
  152. package/dist/__tests__/onboard-preservation.test.js +59 -4
  153. package/dist/__tests__/onboard-preservation.test.js.map +1 -1
  154. package/dist/__tests__/optimize.test.js +2 -0
  155. package/dist/__tests__/optimize.test.js.map +1 -1
  156. package/dist/__tests__/pr-status.test.js +213 -3
  157. package/dist/__tests__/pr-status.test.js.map +1 -1
  158. package/dist/__tests__/pr-workflow-state.test.js +247 -7
  159. package/dist/__tests__/pr-workflow-state.test.js.map +1 -1
  160. package/dist/__tests__/repo-picker.test.js +7 -1
  161. package/dist/__tests__/repo-picker.test.js.map +1 -1
  162. package/dist/__tests__/repository-guidance.test.d.ts +2 -0
  163. package/dist/__tests__/repository-guidance.test.d.ts.map +1 -0
  164. package/dist/__tests__/repository-guidance.test.js +107 -0
  165. package/dist/__tests__/repository-guidance.test.js.map +1 -0
  166. package/dist/__tests__/review-comment-body.test.js +35 -0
  167. package/dist/__tests__/review-comment-body.test.js.map +1 -1
  168. package/dist/__tests__/review-models.test.js +19 -3
  169. package/dist/__tests__/review-models.test.js.map +1 -1
  170. package/dist/__tests__/review-strategy.test.d.ts +2 -0
  171. package/dist/__tests__/review-strategy.test.d.ts.map +1 -0
  172. package/dist/__tests__/review-strategy.test.js +397 -0
  173. package/dist/__tests__/review-strategy.test.js.map +1 -0
  174. package/dist/__tests__/reviewed-sha.test.d.ts +2 -0
  175. package/dist/__tests__/reviewed-sha.test.d.ts.map +1 -0
  176. package/dist/__tests__/reviewed-sha.test.js +116 -0
  177. package/dist/__tests__/reviewed-sha.test.js.map +1 -0
  178. package/dist/__tests__/runner.test.js +206 -2
  179. package/dist/__tests__/runner.test.js.map +1 -1
  180. package/dist/__tests__/skill-attribution.test.d.ts +2 -0
  181. package/dist/__tests__/skill-attribution.test.d.ts.map +1 -0
  182. package/dist/__tests__/skill-attribution.test.js +53 -0
  183. package/dist/__tests__/skill-attribution.test.js.map +1 -0
  184. package/dist/__tests__/skill-broker.test.d.ts +2 -0
  185. package/dist/__tests__/skill-broker.test.d.ts.map +1 -0
  186. package/dist/__tests__/skill-broker.test.js +186 -0
  187. package/dist/__tests__/skill-broker.test.js.map +1 -0
  188. package/dist/__tests__/skill-catalog.test.d.ts +2 -0
  189. package/dist/__tests__/skill-catalog.test.d.ts.map +1 -0
  190. package/dist/__tests__/skill-catalog.test.js +40 -0
  191. package/dist/__tests__/skill-catalog.test.js.map +1 -0
  192. package/dist/__tests__/skill-installer.test.d.ts +2 -0
  193. package/dist/__tests__/skill-installer.test.d.ts.map +1 -0
  194. package/dist/__tests__/skill-installer.test.js +96 -0
  195. package/dist/__tests__/skill-installer.test.js.map +1 -0
  196. package/dist/__tests__/skills-config.test.d.ts +2 -0
  197. package/dist/__tests__/skills-config.test.d.ts.map +1 -0
  198. package/dist/__tests__/skills-config.test.js +12 -0
  199. package/dist/__tests__/skills-config.test.js.map +1 -0
  200. package/dist/__tests__/superseded-fix-pr.test.d.ts +2 -0
  201. package/dist/__tests__/superseded-fix-pr.test.d.ts.map +1 -0
  202. package/dist/__tests__/superseded-fix-pr.test.js +225 -0
  203. package/dist/__tests__/superseded-fix-pr.test.js.map +1 -0
  204. package/dist/__tests__/webhook.test.js +60 -0
  205. package/dist/__tests__/webhook.test.js.map +1 -1
  206. package/dist/cli.js +37 -0
  207. package/dist/cli.js.map +1 -1
  208. package/dist/commands/adoption.d.ts +6 -0
  209. package/dist/commands/adoption.d.ts.map +1 -0
  210. package/dist/commands/adoption.js +132 -0
  211. package/dist/commands/adoption.js.map +1 -0
  212. package/dist/commands/detect-step.d.ts.map +1 -1
  213. package/dist/commands/detect-step.js +5 -1
  214. package/dist/commands/detect-step.js.map +1 -1
  215. package/dist/commands/diagnose.d.ts.map +1 -1
  216. package/dist/commands/diagnose.js +2 -2
  217. package/dist/commands/diagnose.js.map +1 -1
  218. package/dist/commands/init.d.ts.map +1 -1
  219. package/dist/commands/init.js +1 -2
  220. package/dist/commands/init.js.map +1 -1
  221. package/dist/commands/issue.js +1 -1
  222. package/dist/commands/issue.js.map +1 -1
  223. package/dist/commands/kickass.d.ts +1 -1
  224. package/dist/commands/kickass.d.ts.map +1 -1
  225. package/dist/commands/kickass.js +36 -22
  226. package/dist/commands/kickass.js.map +1 -1
  227. package/dist/commands/linear-test.d.ts +18 -0
  228. package/dist/commands/linear-test.d.ts.map +1 -0
  229. package/dist/commands/linear-test.js +130 -0
  230. package/dist/commands/linear-test.js.map +1 -0
  231. package/dist/commands/onboard.d.ts +36 -3
  232. package/dist/commands/onboard.d.ts.map +1 -1
  233. package/dist/commands/onboard.js +263 -45
  234. package/dist/commands/onboard.js.map +1 -1
  235. package/dist/commands/optimize.d.ts.map +1 -1
  236. package/dist/commands/optimize.js +1 -0
  237. package/dist/commands/optimize.js.map +1 -1
  238. package/dist/commands/review.d.ts.map +1 -1
  239. package/dist/commands/review.js +65 -6
  240. package/dist/commands/review.js.map +1 -1
  241. package/dist/commands/run.d.ts.map +1 -1
  242. package/dist/commands/run.js +85 -11
  243. package/dist/commands/run.js.map +1 -1
  244. package/dist/commands/scan.d.ts.map +1 -1
  245. package/dist/commands/scan.js +6 -0
  246. package/dist/commands/scan.js.map +1 -1
  247. package/dist/commands/skill.d.ts +2 -0
  248. package/dist/commands/skill.d.ts.map +1 -0
  249. package/dist/commands/skill.js +16 -0
  250. package/dist/commands/skill.js.map +1 -0
  251. package/dist/commands/status.d.ts.map +1 -1
  252. package/dist/commands/status.js +53 -1
  253. package/dist/commands/status.js.map +1 -1
  254. package/dist/commands/watch.d.ts.map +1 -1
  255. package/dist/commands/watch.js +215 -66
  256. package/dist/commands/watch.js.map +1 -1
  257. package/dist/config/loader.d.ts +3 -1
  258. package/dist/config/loader.d.ts.map +1 -1
  259. package/dist/config/loader.js +19 -6
  260. package/dist/config/loader.js.map +1 -1
  261. package/dist/config/review-model-tiers.json +3 -3
  262. package/dist/config/review-strategy.json +204 -0
  263. package/dist/config/schema.d.ts +269 -15
  264. package/dist/config/schema.d.ts.map +1 -1
  265. package/dist/config/schema.js +109 -10
  266. package/dist/config/schema.js.map +1 -1
  267. package/dist/github/client.d.ts +21 -1
  268. package/dist/github/client.d.ts.map +1 -1
  269. package/dist/github/client.js +46 -7
  270. package/dist/github/client.js.map +1 -1
  271. package/dist/github/detector.d.ts.map +1 -1
  272. package/dist/github/detector.js +28 -1
  273. package/dist/github/detector.js.map +1 -1
  274. package/dist/github/merge.d.ts +0 -5
  275. package/dist/github/merge.d.ts.map +1 -1
  276. package/dist/github/merge.js +0 -9
  277. package/dist/github/merge.js.map +1 -1
  278. package/dist/github/reviewed-sha.d.ts +13 -0
  279. package/dist/github/reviewed-sha.d.ts.map +1 -0
  280. package/dist/github/reviewed-sha.js +62 -0
  281. package/dist/github/reviewed-sha.js.map +1 -0
  282. package/dist/github/superseded-fix-pr.d.ts +19 -0
  283. package/dist/github/superseded-fix-pr.d.ts.map +1 -0
  284. package/dist/github/superseded-fix-pr.js +97 -0
  285. package/dist/github/superseded-fix-pr.js.map +1 -0
  286. package/dist/github/webhook.d.ts +8 -1
  287. package/dist/github/webhook.d.ts.map +1 -1
  288. package/dist/github/webhook.js +8 -1
  289. package/dist/github/webhook.js.map +1 -1
  290. package/dist/issues/ticket-ref.d.ts.map +1 -1
  291. package/dist/issues/ticket-ref.js +6 -5
  292. package/dist/issues/ticket-ref.js.map +1 -1
  293. package/dist/lib/adoption.d.ts +64 -0
  294. package/dist/lib/adoption.d.ts.map +1 -0
  295. package/dist/lib/adoption.js +165 -0
  296. package/dist/lib/adoption.js.map +1 -0
  297. package/dist/lib/annotation.d.ts +7 -0
  298. package/dist/lib/annotation.d.ts.map +1 -1
  299. package/dist/lib/annotation.js +11 -8
  300. package/dist/lib/annotation.js.map +1 -1
  301. package/dist/lib/auto-fix-branch.d.ts +35 -0
  302. package/dist/lib/auto-fix-branch.d.ts.map +1 -0
  303. package/dist/lib/auto-fix-branch.js +88 -0
  304. package/dist/lib/auto-fix-branch.js.map +1 -0
  305. package/dist/lib/board.d.ts +3 -0
  306. package/dist/lib/board.d.ts.map +1 -1
  307. package/dist/lib/board.js +4 -2
  308. package/dist/lib/board.js.map +1 -1
  309. package/dist/lib/clone.d.ts +2 -0
  310. package/dist/lib/clone.d.ts.map +1 -1
  311. package/dist/lib/clone.js +69 -10
  312. package/dist/lib/clone.js.map +1 -1
  313. package/dist/lib/comment-bodies.d.ts +37 -0
  314. package/dist/lib/comment-bodies.d.ts.map +1 -1
  315. package/dist/lib/comment-bodies.js +50 -9
  316. package/dist/lib/comment-bodies.js.map +1 -1
  317. package/dist/lib/logger.d.ts.map +1 -1
  318. package/dist/lib/logger.js +17 -1
  319. package/dist/lib/logger.js.map +1 -1
  320. package/dist/lib/pr-picker.d.ts +1 -1
  321. package/dist/lib/pr-picker.d.ts.map +1 -1
  322. package/dist/lib/pr-picker.js +10 -5
  323. package/dist/lib/pr-picker.js.map +1 -1
  324. package/dist/lib/pr-status.d.ts +2 -1
  325. package/dist/lib/pr-status.d.ts.map +1 -1
  326. package/dist/lib/pr-status.js +70 -6
  327. package/dist/lib/pr-status.js.map +1 -1
  328. package/dist/lib/pr-workflow-state.d.ts +17 -1
  329. package/dist/lib/pr-workflow-state.d.ts.map +1 -1
  330. package/dist/lib/pr-workflow-state.js +110 -5
  331. package/dist/lib/pr-workflow-state.js.map +1 -1
  332. package/dist/lib/repo-picker.d.ts +3 -0
  333. package/dist/lib/repo-picker.d.ts.map +1 -1
  334. package/dist/lib/repo-picker.js +45 -10
  335. package/dist/lib/repo-picker.js.map +1 -1
  336. package/dist/lib/repository-guidance.d.ts +2 -0
  337. package/dist/lib/repository-guidance.d.ts.map +1 -0
  338. package/dist/lib/repository-guidance.js +55 -0
  339. package/dist/lib/repository-guidance.js.map +1 -0
  340. package/dist/lib/review-models.d.ts +15 -2
  341. package/dist/lib/review-models.d.ts.map +1 -1
  342. package/dist/lib/review-models.js +26 -6
  343. package/dist/lib/review-models.js.map +1 -1
  344. package/dist/lib/review-strategy.d.ts +92 -0
  345. package/dist/lib/review-strategy.d.ts.map +1 -0
  346. package/dist/lib/review-strategy.js +282 -0
  347. package/dist/lib/review-strategy.js.map +1 -0
  348. package/dist/lib/runner.d.ts +123 -1
  349. package/dist/lib/runner.d.ts.map +1 -1
  350. package/dist/lib/runner.js +919 -101
  351. package/dist/lib/runner.js.map +1 -1
  352. package/dist/lib/vendor.d.ts +1 -0
  353. package/dist/lib/vendor.d.ts.map +1 -1
  354. package/dist/lib/vendor.js +6 -0
  355. package/dist/lib/vendor.js.map +1 -1
  356. package/dist/lib/workflow.d.ts +13 -0
  357. package/dist/lib/workflow.d.ts.map +1 -1
  358. package/dist/lib/workflow.js +24 -1
  359. package/dist/lib/workflow.js.map +1 -1
  360. package/dist/linear/client.d.ts +18 -0
  361. package/dist/linear/client.d.ts.map +1 -0
  362. package/dist/linear/client.js +67 -0
  363. package/dist/linear/client.js.map +1 -0
  364. package/dist/linear/comment.d.ts +20 -0
  365. package/dist/linear/comment.d.ts.map +1 -0
  366. package/dist/linear/comment.js +57 -0
  367. package/dist/linear/comment.js.map +1 -0
  368. package/dist/linear/identity.d.ts +59 -0
  369. package/dist/linear/identity.d.ts.map +1 -0
  370. package/dist/linear/identity.js +187 -0
  371. package/dist/linear/identity.js.map +1 -0
  372. package/dist/linear/notify.d.ts +35 -0
  373. package/dist/linear/notify.d.ts.map +1 -0
  374. package/dist/linear/notify.js +76 -0
  375. package/dist/linear/notify.js.map +1 -0
  376. package/dist/linear/ref.d.ts +13 -0
  377. package/dist/linear/ref.d.ts.map +1 -0
  378. package/dist/linear/ref.js +90 -0
  379. package/dist/linear/ref.js.map +1 -0
  380. package/dist/linear/verify.d.ts +26 -0
  381. package/dist/linear/verify.d.ts.map +1 -0
  382. package/dist/linear/verify.js +67 -0
  383. package/dist/linear/verify.js.map +1 -0
  384. package/dist/reviewers/address.d.ts +1 -1
  385. package/dist/reviewers/address.d.ts.map +1 -1
  386. package/dist/reviewers/address.js +1 -1
  387. package/dist/reviewers/address.js.map +1 -1
  388. package/dist/reviewers/claude.d.ts +4 -1
  389. package/dist/reviewers/claude.d.ts.map +1 -1
  390. package/dist/reviewers/claude.js +39 -7
  391. package/dist/reviewers/claude.js.map +1 -1
  392. package/dist/reviewers/codex-env.d.ts +12 -0
  393. package/dist/reviewers/codex-env.d.ts.map +1 -0
  394. package/dist/reviewers/codex-env.js +54 -0
  395. package/dist/reviewers/codex-env.js.map +1 -0
  396. package/dist/reviewers/codex.d.ts +18 -1
  397. package/dist/reviewers/codex.d.ts.map +1 -1
  398. package/dist/reviewers/codex.js +164 -81
  399. package/dist/reviewers/codex.js.map +1 -1
  400. package/dist/reviewers/conflict-resolve.d.ts +3 -1
  401. package/dist/reviewers/conflict-resolve.d.ts.map +1 -1
  402. package/dist/reviewers/conflict-resolve.js +21 -6
  403. package/dist/reviewers/conflict-resolve.js.map +1 -1
  404. package/dist/reviewers/fix.d.ts +5 -2
  405. package/dist/reviewers/fix.d.ts.map +1 -1
  406. package/dist/reviewers/fix.js +38 -12
  407. package/dist/reviewers/fix.js.map +1 -1
  408. package/dist/skills/attribution.d.ts +4 -0
  409. package/dist/skills/attribution.d.ts.map +1 -0
  410. package/dist/skills/attribution.js +14 -0
  411. package/dist/skills/attribution.js.map +1 -0
  412. package/dist/skills/broker-server.d.ts +2 -0
  413. package/dist/skills/broker-server.d.ts.map +1 -0
  414. package/dist/skills/broker-server.js +17 -0
  415. package/dist/skills/broker-server.js.map +1 -0
  416. package/dist/skills/broker.d.ts +43 -0
  417. package/dist/skills/broker.d.ts.map +1 -0
  418. package/dist/skills/broker.js +311 -0
  419. package/dist/skills/broker.js.map +1 -0
  420. package/dist/skills/catalog.d.ts +28 -0
  421. package/dist/skills/catalog.d.ts.map +1 -0
  422. package/dist/skills/catalog.js +104 -0
  423. package/dist/skills/catalog.js.map +1 -0
  424. package/dist/skills/installer.d.ts +10 -0
  425. package/dist/skills/installer.d.ts.map +1 -0
  426. package/dist/skills/installer.js +138 -0
  427. package/dist/skills/installer.js.map +1 -0
  428. package/dist/skills/integrity.d.ts +4 -0
  429. package/dist/skills/integrity.d.ts.map +1 -0
  430. package/dist/skills/integrity.js +36 -0
  431. package/dist/skills/integrity.js.map +1 -0
  432. package/docs/dynamic-thoroughness.md +738 -0
  433. package/docs/linear-identity-contract.md +139 -0
  434. package/docs/linear-identity.md +293 -0
  435. package/docs/metrics.md +115 -0
  436. package/docs/trust.md +159 -0
  437. package/get-started.md +322 -16
  438. package/package.json +6 -3
package/README.md CHANGED
@@ -6,509 +6,313 @@
6
6
  <img src="./assets/logo.png" alt="crosscheck" width="160" />
7
7
  </p>
8
8
 
9
- <p align="center"><em>A Humanbased project, built with crosscheck.</em></p>
9
+ <h1 align="center">crosscheck</h1>
10
10
 
11
- # crosscheck
11
+ <p align="center"><strong>Your agents ship fast. Crosscheck makes sure they ship right.</strong></p>
12
12
 
13
13
  <p align="center">
14
- <img src="./assets/screenshot-watch.png" alt="crosscheck watch — live pipeline view" width="860" />
14
+ <a href="https://www.npmjs.com/package/@humanbased/crosscheck"><img src="https://img.shields.io/npm/v/@humanbased/crosscheck?color=2f6feb&label=npm" alt="npm" /></a>
15
+ <a href="./LICENSE"><img src="https://img.shields.io/badge/license-MIT-blue" alt="MIT" /></a>
16
+ <a href="https://nodejs.org"><img src="https://img.shields.io/badge/node-18%2B-brightgreen" alt="Node 18+" /></a>
15
17
  </p>
16
18
 
17
- **Your agents ship fast. Crosscheck makes sure they ship right.**
18
-
19
- AI coding agents create PRs faster than review habits can absorb. The failure mode isn't broken builds — it's *early victory*: patches that pass CI, look complete, and still hide regressions, brittle edge cases, or half-finished fixes.
20
-
21
- Crosscheck adds an independent Review → Fix → Recheck loop. One agent writes the patch. Another reviews it. Findings go back to the author to repair. The result gets rechecked before merge. The PR moves toward genuinely merge-ready — not just "looks green."
22
-
23
- No new hosted service. No per-review API bill. Crosscheck runs through the `claude` and `codex` CLIs you already have — your existing subscriptions, your machine or server.
24
-
25
- Built by [Humanbased](https://github.com/humanbased-ai). Read the field report: [What 295 Agentic PRs Taught Us About Code Review](https://blog.humanbased.ai/posts/agentic-pr-quality-crosscheck/) — 295 agentic PRs analyzed, real Crosscheck logs included.
26
-
27
- ## Why crosscheck?
28
-
29
- **Agent velocity without lowering the merge bar.**
30
-
31
- - **Independent eyes, not self-review** — route Claude-authored PRs to Codex and vice versa. Self-review is exactly where early-victory failures hide.
32
- - **Review → Fix → Recheck, not just comments** — findings return to the author agent for repair; a clean recheck follows before merge. PRs move forward, not sideways.
33
- - **No new vendor** — runs through the `claude` and `codex` CLIs you already pay for. No per-review bill, no extra trust surface.
34
- - **Configurable for any team size** — review-only, review + fix, or the full loop. Use one workflow locally, or one always-on team watcher with per-repo overrides via `crosscheck alter`.
35
-
36
- ## Who uses crosscheck
37
-
38
- | Persona | Problem | How crosscheck helps |
39
- |---|---|---|
40
- | **Solo agentic builder** | Same agent that wrote the code may self-approve incomplete work | Independent reviewer from a different vendor, on your machine |
41
- | **Technical founder** | AI PRs look done before delivering stable value | Closes the loop: review finding → agent fix → clean recheck |
42
- | **Engineering lead** | Agent use is hard to supervise or standardize | A default full-loop workflow, per-repo overrides (`crosscheck alter`), and a visible PR audit trail |
43
- | **OSS maintainer** | Review bandwidth is scarce; comments must be actionable | One-shot `crosscheck review` posts concrete findings directly on the PR |
19
+ <p align="center">
20
+ <img src="./assets/screenshot-watch.png" alt="crosscheck watch — live pipeline view" width="860" />
21
+ </p>
44
22
 
45
- ### Usage scenarios
23
+ ---
46
24
 
47
- **Local use**
25
+ ## The problem
48
26
 
49
- Use Crosscheck from your own machine when you want an independent review before merge, or a temporary watcher while you work.
27
+ AI coding agents open PRs faster than review habits can absorb them. The failure mode isn't a broken build — it's **early victory**: a patch that passes CI, reads as complete, and quietly carries a regression, a brittle edge case, or a half-finished fix.
50
28
 
51
- ```bash
52
- # Catch regressions before merging a solo PR
53
- crosscheck run <pr-url>
29
+ Asking the agent that wrote the patch to review it doesn't help. That's exactly where early victory hides.
54
30
 
55
- # One-shot review of a specific PR
56
- crosscheck review <pr-url>
31
+ ## What crosscheck does
57
32
 
58
- # Continuous review while your terminal is open
59
- crosscheck onboard --personal
60
- crosscheck watch
33
+ One agent writes the patch. **A different one reviews it.** Findings go back to the author agent to repair, and the result is rechecked before merge.
61
34
 
62
- # Loop until the agent produces an approved patch
63
- crosscheck run <pr-url> --crazy
64
35
  ```
65
-
66
- **Always-on team server**
67
-
68
- Run one long-lived `watch` process for the whole team or org on a shared machine. Configure the default workflow once, then narrow individual repos only when needed.
69
-
70
- ```bash
71
- # One-time setup on the server
72
- crosscheck onboard --team
73
-
74
- # Example: one repo gets review-only, the rest keep the global workflow
75
- crosscheck alter humanbased-ai/xny-monorepo --review-only
76
-
77
- # Start the team watcher
78
- crosscheck watch
36
+ PR → review → fix → recheck → merge-ready
37
+ (codex) (claude) (codex)
79
38
  ```
80
39
 
81
- ---
40
+ Three properties make that practical:
82
41
 
83
- ## Quick start
42
+ - **Independent eyes.** Claude-authored PRs route to Codex and vice versa. Origin is detected from the PR body, commit trailers, and branch prefix — no manual tagging.
43
+ - **A loop, not a comment.** Findings return to the author agent for repair; a clean recheck follows. The PR moves forward instead of sideways.
44
+ - **No new vendor.** Runs through the `claude` and `codex` CLIs you already pay for. No hosted service, no per-review API bill, no extra trust surface.
84
45
 
85
- ### First useful review in 10 minutes
46
+ ## Same mission. Sharper skills.
86
47
 
87
- Start with one low-risk PR before turning on continuous watch mode. You only need GitHub CLI plus one authenticated reviewer CLI.
48
+ Crosscheck stays focused on one job: making agent-authored PRs trustworthy. It now ships supercharged with preloaded, coding-specialized skills that every invoked coding agent can use during review, diagnosis, repair, recheck, and conflict resolution.
88
49
 
89
- ```bash
90
- # 1. Install crosscheck
91
- npm install -g @humanbased/crosscheck
50
+ The recommended onboarding bundle combines a broad review baseline, architecture vocabulary, and rigorous bug diagnosis:
92
51
 
93
- # 2. Authenticate GitHub
94
- brew install gh && gh auth login
52
+ - `code-review-skill (by @awesome-skills, MIT)` — comprehensive review guidance across languages, architecture, security, and performance.
53
+ - `codebase-design (by @mattpocock, MIT)` evaluates deep modules, small interfaces, clean seams, and testability.
54
+ - `diagnosing-bugs (by @mattpocock, MIT)` — requires a reproducible signal, tested hypotheses, and regression evidence before declaring a fix complete.
95
55
 
96
- # 3. Authenticate one reviewer
97
- npm install -g @openai/codex && codex login --device-auth
98
- # or:
99
- npm install -g @anthropic-ai/claude-code && claude
56
+ Matt Pocock's `code-review` is also preloaded as an alternative for evidence-rich repositories with documented standards and a clear issue or PRD. It remains off by default because it competes with the broad `code-review-skill`; onboarding warns instead of loading both. Enable only the practices your team wants, or install your own skill with `crosscheck skill install <source>`.
100
57
 
101
- # 4. Check your setup
102
- crosscheck status
58
+ Enabled skills are available, not blindly forced: the coding agent decides which are relevant to each operation. The terminal and PR comment then attribute only the skills actually activated for that step. Crosscheck also honors the target repository's base-branch `AGENTS.md` and `CLAUDE.md` guidance; those local practices take precedence over bundled skill advice.
103
59
 
104
- # 5. Review the public fixture PR
105
- crosscheck review https://github.com/humanbased-ai/crosscheck-proof-fixture/pull/1 --reviewer codex
106
- ```
60
+ Setup and trust model: **[Agent skills](./get-started.md#crosscheck-skill-install-source)**.
107
61
 
108
- This fixture PR intentionally contains a realistic agentic-code regression, so you can see whether Crosscheck produces a useful review before pointing it at your own repo. Use `--reviewer claude` if Claude Code is the authenticated reviewer. After the fixture review works, swap in one low-risk PR from your repo, then run `crosscheck onboard` to configure repos, workflow mode, and continuous monitoring.
62
+ Built by [Humanbased](https://github.com/humanbased-ai). Field report: [What 295 Agentic PRs Taught Us About Code Review](https://blog.humanbased.ai/posts/agentic-pr-quality-crosscheck/).
63
+
64
+ ---
109
65
 
110
- ### Continuous local mode
66
+ ## Install
111
67
 
112
68
  ```bash
113
- # 1. Install crosscheck and the agent CLIs
114
69
  npm install -g @humanbased/crosscheck
115
- npm install -g @anthropic-ai/claude-code && claude # Claude Pro/Max subscription
116
- npm install -g @openai/codex && codex login --device-auth # ChatGPT Plus/Pro subscription
117
- brew install gh && gh auth login # GitHub CLI
118
-
119
- # 2. Guided setup — repos, review mode, workflow pipeline
120
- crosscheck onboard
121
-
122
- # 3. Start watching
123
- crosscheck watch # continuous review → fix → recheck as PRs arrive
124
70
  ```
125
71
 
126
- > Want reviews only (no auto-fix) for a repo? Make it review-only with
127
- > `crosscheck alter owner/repo --review-only` — see [per-repo overrides](#crosscheck-alter-repo).
128
-
129
- ### Always-on team mode
72
+ <details>
73
+ <summary>Other channels</summary>
130
74
 
131
75
  ```bash
132
- # 1. Run guided setup on the shared machine
133
- crosscheck onboard --team
76
+ npm install -g @humanbased/crosscheck@beta # latest features, rougher edges
77
+ npx @humanbased/crosscheck <command> # no install
134
78
 
135
- # 2. Optional: tune individual repos without changing the global workflow
136
- crosscheck alter humanbased-ai/xny-monorepo --review-only
137
- crosscheck alter humanbased-ai/api --steps review,fix,recheck
138
-
139
- # 3. Start the long-lived watcher
140
- crosscheck watch
79
+ git clone https://github.com/humanbased-ai/crosscheck
80
+ cd crosscheck && npm install && npm run build && npm link
141
81
  ```
82
+ </details>
142
83
 
143
- ---
144
-
145
- ## Commands
84
+ You need GitHub CLI plus **at least one** reviewer CLI. Install both only if you want cross-vendor routing.
146
85
 
147
86
  ```bash
148
- crosscheck onboard # guided setup — pick repos, mode, and default pipeline
149
- crosscheck alter <repo> # per-repo override: --steps review,fix | --review-only | --reset | --show
150
- crosscheck watch # continuous use tunnel + webhook + listening
151
- crosscheck review <pr-urls...> # review one or more PRs (comma lists, ranges, cross-repo)
152
- crosscheck run <pr-urls...> # run the full workflow: review → (fix → recheck) × max_rounds (--review-only for review only)
153
- crosscheck recheck|fix|resolve <pr-urls...> # force one workflow step on one or more PRs
154
- crosscheck scan # show open PR workflow state across monitored repos
155
- crosscheck detect-step <pr-url> # explain the next workflow step for one PR
156
- crosscheck kickass # advance stale PRs from an interactive operator queue
157
- crosscheck init # check prerequisites, write starter config
158
- crosscheck status # auth state, config summary, CLI versions
87
+ gh auth login
88
+ npm install -g @anthropic-ai/claude-code && claude # Claude Pro or Max
89
+ npm install -g @openai/codex && codex login --device-auth # ChatGPT Plus or Pro
159
90
  ```
160
91
 
161
- **Operator queue (scan + kickass)**
162
-
163
- `crosscheck scan` tracks two independent dimensions per PR:
164
-
165
- | Workflow stage (`reviewState`) | Meaning | Next action |
166
- |---|---|---|
167
- | `NEEDS_REVIEW` | No crosscheck review for current HEAD | review |
168
- | `NEEDS_FIX` | Reviewed — fix requested | fix |
169
- | `NEEDS_RECHECK` | Fix committed, recheck pending | recheck |
170
- | `APPROVED` | Reviewed and approved | merge |
171
-
172
- | Verdict (`verdict`) | Meaning |
173
- |---|---|
174
- | `UNREVIEWED` | No review found |
175
- | `APPROVE` | AI approved |
176
- | `NEEDS_WORK` | AI requested changes |
177
- | `BLOCK` | AI hard-blocked merge |
92
+ Both reviewers run on your existing subscription — no API key required.
178
93
 
179
- `BLOCK` and `NEEDS_WORK` both map to `NEEDS_FIX` stage — same next action, but the `verdict` field preserves severity so operators can prioritise.
180
-
181
- **How workflow steps are counted**
182
-
183
- Crosscheck reconstructs PR workflow state from visible artifacts:
184
-
185
- | Evidence | Counts as |
186
- |---|---|
187
- | Review or recheck comment with `<!-- crosscheck: ... verdict=... -->` | completed `review` / `recheck` step |
188
- | Fix or conflict-resolve comment, such as `<!-- crosscheck: fix_applied ... -->` | completed `fix` / `conflict-resolve` step |
189
- | PR commit trailer, such as `Crosscheck-Step: fix` | completed step declared by that trailer |
190
-
191
- Commit trailers are accepted as operator-declared workflow state. In practice, a PR author may command Claude, Codex, or another agent to apply a fix outside a standalone Crosscheck post; if the resulting PR commit carries `Crosscheck-Step: fix`, Crosscheck counts it as fix evidence.
192
-
193
- That evidence only advances the next step to `recheck` when the fix commit is the current PR HEAD. If another commit lands after the fix evidence, Crosscheck starts a fresh review round so the newer code is reviewed normally. This prevents an old fix trailer from marking later changes as ready for recheck.
94
+ ## First review in two minutes
194
95
 
195
96
  ```bash
196
- crosscheck scan [--tidy] [--stale-after <duration>] [--force] [--json]
197
- crosscheck kickass [--dry-run] [--stale-after <duration>] [--force]
97
+ crosscheck status # confirm auth
98
+ crosscheck review https://github.com/humanbased-ai/crosscheck-proof-fixture/pull/1 --reviewer codex
198
99
  ```
199
100
 
200
- `crosscheck review --reviewer`, `crosscheck run --reviewer`, `crosscheck run --fixer`, and `crosscheck run --vendor` accept vendor aliases:
201
- - Claude: `claude`, `claude-code`, `cc`, `anthropic`
202
- - Codex: `codex`, `openai`
203
-
204
- **Continuous improvement** *(experimental)*
101
+ That clones the branch, reviews it against base, and posts a comment on the PR. Once the fixture produces a useful verdict, point it at one low-risk PR of your own — then set up continuous review:
205
102
 
206
103
  ```bash
207
- crosscheck diagnose # surface failure patterns from review logs
208
- crosscheck optimize [--apply] # rewrite reviewer instructions based on diagnose output
209
- crosscheck impact [--money] # time saved, issues caught, code quality trends
210
- crosscheck issue # draft and file a bug report from recent error logs
104
+ crosscheck onboard # guided: repos, routing, pipeline depth, connection
105
+ crosscheck watch # listen for PR events
211
106
  ```
212
107
 
213
108
  ---
214
109
 
215
- ### `crosscheck onboard`
110
+ ## Where results land
216
111
 
217
- Interactive setup wizard. Picks repos/orgs to monitor, selects single-vendor or cross-vendor mode, configures the review pipeline, and writes `~/.crosscheck/config.yml` and `workflow.yml`.
112
+ ### On the pull request
218
113
 
219
- ```bash
220
- crosscheck onboard # guided setup
221
- crosscheck onboard --personal # skip persona prompt, go straight to personal mode
222
- crosscheck onboard --team # skip persona prompt, go straight to team mode
223
- crosscheck onboard -y # accept all defaults non-interactively
224
- ```
225
-
226
- ---
227
-
228
- ### `crosscheck alter <repo>`
229
-
230
- Sets the workflow depth for one repo, leaving the global default in place for every other monitored repo. Writes a standalone file at `~/.crosscheck/workflows/<owner>__<repo>.yml` (`alter-workflow` is an alias). This is how you run one watcher for many repos while making one repo review-only. Changes apply on the next PR event — no need to restart `crosscheck watch`.
114
+ Every review posts a comment carrying a machine-readable annotation:
231
115
 
232
- ```bash
233
- crosscheck alter humanbased-ai/xny-monorepo --review-only # alias for --steps review
234
- crosscheck alter github.com/humanbased-ai/xny-monorepo --steps review,fix
235
- crosscheck alter https://github.com/humanbased-ai/xny-monorepo --steps review,fix,recheck
236
- crosscheck alter humanbased-ai/xny-monorepo --show # print effective steps
237
- crosscheck alter humanbased-ai/xny-monorepo --reset # revert to the global default
238
116
  ```
239
-
240
- Accepted repo formats: `owner/repo`, `github.com/owner/repo`, and `https://github.com/owner/repo`. The override narrows the global `~/.crosscheck/workflow.yml` it wins over the global default but a repo-committed `.crosscheck/workflow.yml` still wins over it.
241
-
242
- ---
243
-
244
- ### `crosscheck watch`
245
-
246
- Starts an SSH tunnel (localhost.run), registers GitHub webhooks, and listens for PR events. Everything self-cleans on Ctrl+C.
247
-
248
- ```bash
249
- crosscheck watch
250
- crosscheck watch --no-backtrace # skip startup scan for unreviewed open PRs
251
- crosscheck watch --reconfigure # re-run deployment setup before starting
117
+ <!-- crosscheck: origin=claude reviewer=codex model=gpt-5.6-terra
118
+ type=review round=1 verdict=NEEDS_WORK service=crosscheck sha=a1b2c3d -->
252
119
  ```
253
120
 
254
- > For reviews only, make the repo review-only with `crosscheck alter <repo> --review-only` rather than a global flag.
121
+ That tag is the audit trail. It's how crosscheck knows which step ran, what verdict came back, and what to do next and it's a stable contract you can parse.
255
122
 
256
- ---
123
+ ### On your Linear issue
257
124
 
258
- ### `crosscheck review <pr-urls...>`
125
+ Optional, off by default. When enabled, the verdict is mirrored onto the Linear issue the PR belongs to, so outcomes show up where work is planned:
259
126
 
260
- Reviews one or more PRs. Clones, checks out, reviews, and posts the comment. The PR argument accepts the [multi-PR spec syntax](#multi-pr-syntax) — multiple PRs are reviewed concurrently.
127
+ <p align="center">
128
+ <img src="./assets/linear-comment.svg" alt="A crosscheck review comment on a Linear issue" width="740" />
129
+ </p>
261
130
 
262
- ```bash
263
- crosscheck review https://github.com/org/repo/pull/42
264
- crosscheck review <pr-url> --reviewer claude # force Claude regardless of detection
265
- crosscheck review <pr-url> --reviewer codex # force Codex regardless of detection
266
- crosscheck review <pr-url> --reviewer cc # alias for Claude
267
- crosscheck review <pr-url> --reviewer openai # alias for Codex
268
- crosscheck review .../pull/245,255 # review several PRs at once
269
- crosscheck review .../pull/245-256 # review an inclusive range
270
- ```
131
+ Attribution is a ladder — **start at the bottom, climb only if you need to**:
271
132
 
272
- ---
133
+ | Rung | Setup | Comments appear as |
134
+ |---|---|---|
135
+ | **api key** | one env var | Your Linear account, with a `🤖 crosscheck · <model>` signature line |
136
+ | **workspace app** | one OAuth app, ~5 min, once per workspace | crosscheck itself, with its own icon |
273
137
 
274
- ### `crosscheck run <pr-urls...>`
138
+ The API key rung is fully functional — it finds the issue and posts the comment. What it lacks is *attribution*, not capability. So the question isn't which is better, it's **how many things write to your workspace**. If you're the only one, the app is ceremony.
275
139
 
276
- Runs the configured workflow against one or more PRs: review → (fix → recheck) × `max_rounds`. Without `--steps`, this honors any repo workflow set by `crosscheck alter`. Loops autonomously through fix→recheck cycles up to the `max_rounds` value configured in `workflow.yml` (default: 1). Use `--crazy` or `--half-crazy` to loop until approved or unblocked, ignoring `max_rounds`. The PR argument accepts the [multi-PR spec syntax](#multi-pr-syntax); multiple PRs run concurrently (one agent per PR by default).
140
+ `crosscheck onboard` asks which rung you want and writes the config:
277
141
 
278
- ```bash
279
- crosscheck run <pr-url>
280
- crosscheck run <pr-url> --steps review # only the review step
281
- crosscheck run <pr-url> --steps fix,recheck # skip initial review
282
- crosscheck run <pr-url> --reviewer claude # force review/recheck agent
283
- crosscheck run <pr-url> --fixer claude # force fix agent
284
- crosscheck run <pr-url> --vendor claude # force review/recheck/fix agent
285
- crosscheck run <pr-url> --dry-run # review without posting or fixing
286
- crosscheck run <pr-url> --crazy # 🔥🔥 loop until APPROVE
287
- crosscheck run <pr-url> --half-crazy # 🔥 loop until not BLOCK
288
- crosscheck run <pr-url> --timeout 10m # custom reviewer timeout
289
- crosscheck run .../pull/245,255 # several PRs, concurrently
290
- crosscheck run .../pull/245-256 --concurrent 3 # range, max 3 agents in parallel
291
- ```
142
+ <p align="center">
143
+ <img src="./assets/linear-onboard.svg" alt="crosscheck onboard — choosing a Linear attribution rung" width="700" />
144
+ </p>
292
145
 
293
- ---
146
+ To check a setup without waiting for a PR, `linear-test` runs the whole path and posts nothing:
294
147
 
295
- ### `crosscheck recheck` / `fix` / `resolve <pr-urls...>`
148
+ <p align="center">
149
+ <img src="./assets/linear-test.svg" alt="crosscheck linear-test — verifying Linear write-back end to end" width="700" />
150
+ </p>
296
151
 
297
- Force a single workflow step against one or more PRs, bypassing next-step auto-detection. Each is sugar for `crosscheck run <spec> --steps <type>` and accepts the same [multi-PR spec syntax](#multi-pr-syntax) and `--concurrent` / `--sequential` / `--stagger` flags as `run`.
152
+ <details>
153
+ <summary>Confirming which rung you're on at any time</summary>
298
154
 
299
- | Command | Forces the step |
300
- |---|---|
301
- | `crosscheck recheck <spec>` | `recheck` — re-evaluate against the latest review |
302
- | `crosscheck fix <spec>` | `fix` — apply fixes for the latest review |
303
- | `crosscheck resolve <spec>` | `conflict-resolve` — resolve merge conflicts (Claude only) |
155
+ `crosscheck status` resolves the configured identity for real and reports what a write would render as:
304
156
 
305
- ```bash
306
- crosscheck recheck https://github.com/org/repo/pull/42
307
- crosscheck fix .../pull/245,255 --fixer claude
308
- crosscheck resolve .../pull/245-256 --vendor claude
309
- ```
157
+ <p align="center">
158
+ <img src="./assets/linear-status.svg" alt="crosscheck status — the Linear identity section" width="620" />
159
+ </p>
160
+ </details>
310
161
 
311
- When the step is absent from the active `workflow.yml`, `recheck` and `conflict-resolve` are synthesized with built-in defaults so the command still runs.
162
+ Full walkthrough: **[docs/linear-identity.md](./docs/linear-identity.md)**.
312
163
 
313
164
  ---
314
165
 
315
- ### Multi-PR syntax
166
+ ## Commands
316
167
 
317
- `run`, `review`, `recheck`, `fix`, and `resolve` accept a single PR URL or a **spec** that expands to many PRs — a comma-separated list of full URLs, bare numbers, and `N-M` ranges:
168
+ | Command | What it does |
169
+ |---|---|
170
+ | `crosscheck onboard` | Guided setup — repos, routing, pipeline depth, connection |
171
+ | `crosscheck status` | Auth, config, Linear identity, logs, impact summary |
172
+ | `crosscheck skill install <source>` | Install an Agent Skill from Git or a local directory |
173
+ | `crosscheck review <pr>` | One-shot review, posts a comment |
174
+ | `crosscheck run <pr>` | Full pipeline for a PR — review, fix, recheck |
175
+ | `crosscheck recheck` / `fix` / `resolve` | Run one step in isolation |
176
+ | `crosscheck watch` | Listen for PR events and run the pipeline automatically |
177
+ | `crosscheck scan` | Show open PRs with stale crosscheck state |
178
+ | `crosscheck kickass` | Pick a stale PR and drive it to its next step |
179
+ | `crosscheck alter <repo>` | Set a per-repo pipeline depth |
180
+ | `crosscheck detect-step <pr>` | Show step history and the next step to run |
181
+ | `crosscheck linear-test [issue]` | Dry-run Linear write-back |
182
+ | `crosscheck diagnose` / `optimize` / `impact` / `adoption` / `issue` | Analyse logs, tune config, report value, report usage, file tickets |
183
+
184
+ Multi-PR forms work where sensible — comma lists, bare numbers, and ranges:
318
185
 
319
186
  ```bash
320
- .../pull/245,255 # two PRs in the same repo
321
- .../pull/245-256 # an inclusive range
322
- .../repo/pull/245,https://github.com/o/other/pull/3 # across repos
187
+ crosscheck review https://github.com/acme/app/pull/245,255
188
+ crosscheck run https://github.com/acme/app/pull/245-256 --concurrent 4
323
189
  ```
324
190
 
325
- The first token must be a full URL (a bare number inherits the most recent repo). Duplicates are de-duplicated and a spec expands to at most 100 PRs. Multiple PRs run concurrently by default — control parallelism with `--concurrent <n>`, `--sequential`, or `--stagger <ms>`.
191
+ Full flag reference: **[get-started.md](./get-started.md)**.
326
192
 
327
193
  ---
328
194
 
329
- ### `crosscheck scan`
195
+ ## Configuration
330
196
 
331
- Scans every open PR in the configured monitor scope and reports where each one is in the crosscheck workflow. Results are cached for 60 seconds.
197
+ Config lives at `~/.crosscheck/config.yml`. A `./crosscheck.config.yml` in the working directory is treated as a deliberate per-project override.
332
198
 
333
- States: `NEEDS_REVIEW` · `NEEDS_FIX` · `BLOCK` · `NEEDS_RECHECK` · `APPROVE`
199
+ ### Review depth
334
200
 
335
- ```bash
336
- crosscheck scan # all open PRs, grouped stale/not-stale
337
- crosscheck scan --tidy # stale actionable rows only
338
- crosscheck scan --stale-after 4h # custom staleness threshold (default 24h)
339
- crosscheck scan --force # bypass cache
340
- crosscheck scan --json # machine-readable output
201
+ ```yaml
202
+ quality:
203
+ mode: smart # smart (default) | fixed
204
+ tier: balanced # fast | balanced | thorough the fallback under smart
205
+
206
+ skills:
207
+ enabled:
208
+ - code-review-skill # recommended · @awesome-skills, MIT
209
+ - codebase-design # recommended · @mattpocock, MIT
210
+ - diagnosing-bugs # recommended · @mattpocock, MIT
341
211
  ```
342
212
 
343
- ---
344
-
345
- ### `crosscheck detect-step`
346
-
347
- Explains the workflow history for one PR and prints the next step Crosscheck would run. Use this when a PR has mixed evidence from comments, Crosscheck commits, or ad hoc agent commits with `Crosscheck-Step` trailers.
213
+ Agents decide whether an enabled skill applies to each review, fix, recheck, or conflict-resolution step. A skill is activated only after the agent successfully loads it through the broker; activation lasts for that step session, including retries, and does not carry into later steps or runs. PR comments preserve only that completed step's activated-skill attribution.
214
+ Existing configs keep skills disabled on upgrade; use `crosscheck onboard` to opt in. Installed packages are integrity-checked before agents can load them.
348
215
 
349
- ```bash
350
- crosscheck detect-step <pr-url>
351
- crosscheck detect-step <pr-url> --json
352
- ```
353
-
354
- ---
216
+ For review and recheck, Crosscheck also applies repository-defined review practices from `AGENTS.md` and `CLAUDE.md`. In monorepos it combines root guidance with the files scoped to changed paths, using the trusted base-branch versions so a PR cannot rewrite its own review rules.
355
217
 
356
- ### `crosscheck kickass`
218
+ | Tier | Claude | Codex | Cost per review¹ |
219
+ |---|---|---|---|
220
+ | `fast` | Haiku 4.5 | GPT-5.6 Luna | $0.24 · $0.06 |
221
+ | `balanced` | Sonnet 5 | GPT-5.6 Terra | $0.72 · $0.58 |
222
+ | `thorough` | **Opus 5** | GPT-5.6 Sol | $1.20 · $1.44 |
357
223
 
358
- Selects stale PRs from the operator queue and advances them runs `scan` first, presents a multi-select picker, shows a preflight summary, then executes after confirmation.
224
+ ¹ Output-token cost at 48k output tokens, the measured median for one review. A review is an agentic session, not a single call expect **10–16 minutes** of wall clock (median 643s, p90 984s across 43 logged runs). Tier changes depth and the subprocess timeout, not seconds-scale latency.
359
225
 
360
- ```bash
361
- crosscheck kickass # interactive operator queue
362
- crosscheck kickass --dry-run # preflight only — no mutations
363
- crosscheck kickass --stale-after 2h # tighter staleness threshold
364
- crosscheck kickass --force # bypass scan cache before picking
365
- crosscheck kickass --crazy # 🔥🔥 auto loop until APPROVE
366
- crosscheck kickass --half-crazy # 🔥 auto loop until not BLOCK
367
- ```
226
+ `claude-fable-5` is **banned** from review: 2× Opus 5's price for a lower coding benchmark score.
368
227
 
369
- Actions: `NEEDS_REVIEW CR` · `NEEDS_FIX/BLOCK → Fix` · `NEEDS_RECHECK → Recheck` · `APPROVE → Merge`
228
+ ### Dynamic thoroughness (`mode: smart`)
370
229
 
371
- **`kickass` + `watch` combo**
230
+ **On by default.** Instead of one tier for every call, Crosscheck classifies each PR from its changed-file list and adjusts model and effort to match. Classification runs on the already-cloned working copy, so it costs one `git diff` and no API call.
372
231
 
373
- For the best recovery experience when a batch of PRs is stuck (timed out, stopped before `watch` was running), run both commands together. Each plays a distinct role:
232
+ | # | PR class | Detected by | Tier | Steps |
233
+ |---|---|---|---|---|
234
+ | 1 | Generated / vendored | every file is a lockfile or build output | — | **PR skipped** |
235
+ | 2 | Security / data-critical | auth, crypto, payment, migration paths; `risk:T3`; hotfix→default | `thorough` | full loop |
236
+ | 3 | Deletion-only | ≤ 5 additions with ≥ 20 deletions | `fast` | review |
237
+ | 4 | Docs / spec | ≥ 50% Markdown | `balanced` | review |
238
+ | 5 | Test-only | every file is a test | `fast` | review, fix |
239
+ | 6 | Config / infra | ≥ 50% config, no source | `balanced` | full loop |
240
+ | 7 | Trivial | ≤ 3 files, ≤ 150 lines | `fast` | review, fix |
241
+ | 8 | Standard | everything else | `balanced` | full loop |
374
242
 
375
- - `kickass` kicks each stuck PR **one step at a time** it uses `detect-step` to read live PR history and dispatches only the next needed step (review, fix, or recheck).
376
- - `watch` owns **all continuation** — it listens for the webhooks each completed step produces and runs the full remaining pipeline from there.
243
+ **Order is the routing logic** first match wins, and security sits second so it dominates every cheapening rule below it. A deletion that removes auth code, or a two-file migration, is never routed to `fast`.
377
244
 
378
- ```
379
- crosscheck kickass
380
- └─ ck run <url> --trigger kickass (one step; detect-step finds where to start)
381
- └─ detect-step → "review" run review only → posts comment
382
- └─ detect-step → "fix" run fix only → pushes commit
383
- └─ detect-step → "recheck" run recheck only → posts verdict
245
+ Classification may set a **floor**, or promote on **consequence** — a security path is reviewed thoroughly because a miss there is expensive. It may **not** predict that a PR will be hard: across a 400-PR census, diff size correlates only 0.51 with realized review cost (the largest one-call PR was 101k lines; the most expensive changed 2 files). So escalation responds to what the review actually found — round 2 raises effort, round 3 switches vendor, then it hands off to a human rather than looping.
384
246
 
385
- crosscheck watch
386
- ├─ issue_comment (type=review) → pick up fix step automatically
387
- └─ synchronize (fix commit) → pick up recheck step automatically
388
- ```
247
+ Class tier, effort, **and** step set are all applied. The class is resolved once per workflow, not per step — the fix step pushes commits, so re-classifying could make the review and recheck comments cite different tiers for the same PR.
389
248
 
390
- > **Note:** `crosscheck run <pr-url>` invoked directly runs the **full remaining pipeline** from the detected starting step. The one-step behaviour above applies only when kickass dispatches it with `--trigger kickass`.
249
+ The step set **narrows** the configured pipeline and never widens it: a repo pinned to review-only with `crosscheck alter` stays review-only whatever the class says.
391
250
 
392
- Start `watch` first, then run `kickass` in a second terminal:
251
+ Rounds beyond the first escalate on measured non-convergence rather than prediction — effort rises where the model supports it, the tier is promoted where it does not, and the model never weakens.
393
252
 
394
- ```bash
395
- # terminal 1
396
- crosscheck watch
253
+ Every comment says which policy produced it:
397
254
 
398
- # terminal 2
399
- crosscheck scan --force # refresh PR state
400
- crosscheck kickass
401
255
  ```
256
+ _thorough tier · touches a security or data-critical path, where a missed defect
257
+ is expensive · strategy v1.1.0_
402
258
 
403
- > **How the review→fix bridge works:** after `kickass` posts a review comment, GitHub fires an `issue_comment` webhook (not a `pull_request` event). `watch` subscribes to `issue_comment` and, when it sees a crosscheck `type=review` annotation on an open PR, fetches the current PR head and runs the fix step automatically — no new commit required to wake it up. (Introduced in [#193](https://github.com/Motivation-Labs/crosscheck/pull/193).)
404
-
405
- **Autonomous loop modes**
406
-
407
- `--crazy` and `--half-crazy` turn `run` and `kickass` into autonomous fix→recheck loops that keep going until the verdict improves — no manual re-runs needed.
408
-
409
- | Flag | Stops when | Max rounds | Timeout |
410
- |---|---|---|---|
411
- | `--crazy` 🔥🔥 | verdict = `APPROVE` | ∞ | none |
412
- | `--half-crazy` 🔥 | verdict ≠ `BLOCK` | ∞ | none |
413
-
414
- Both flags disable all reviewer subprocess timeout constraints — long fixes on large PRs won't be cut short. Use `--timeout <duration>` (e.g. `--timeout 10m`) without these flags to set a custom cap.
415
-
416
- ```bash
417
- # Run full workflow and keep looping until approved
418
- crosscheck run <pr-url> --crazy
419
-
420
- # Advance every stale PR until it's no longer blocked
421
- crosscheck kickass --half-crazy
422
-
423
- # Custom timeout without looping
424
- crosscheck run <pr-url> --timeout 10m
259
+ <!-- crosscheck: model=claude-opus-5 verdict=BLOCK strategy=1.1.0 class=risky tier=thorough -->
425
260
  ```
426
261
 
427
- ---
428
-
429
- ## Configuration
262
+ The rationale is the matched class's own `reason` field, not prose written per review, so the explanation and the routing decision cannot drift apart. A review from six weeks ago stays explicable after the policy changes.
430
263
 
431
- Crosscheck uses `~/.crosscheck/config.yml` by default. If that file exists, it wins over `./crosscheck.config.yml` unless you pass `--config ./crosscheck.config.yml`.
264
+ The policy is versioned in [`src/config/review-strategy.json`](./src/config/review-strategy.json), carries its own sources and a 60-day review interval, and is checked weekly by the [`Review Strategy`](./.github/workflows/review-strategy.yml) workflow — verify it any time with `npm run verify:strategy`. Full evidence: [docs/dynamic-thoroughness.md](./docs/dynamic-thoroughness.md).
432
265
 
433
- ### Review depth (`quality.tier`)
266
+ > **Leave `vendors.*.model` unset under smart mode.** An explicit model outranks the strategy, so pinning one makes per-PR selection a no-op. When that happens crosscheck **withholds** the tier from the comment rather than citing a routing decision that did not happen. `crosscheck onboard` clears the pin — and prints what it cleared — when you choose smart.
434
267
 
435
- ```yaml
436
- # crosscheck.config.yml
437
- quality:
438
- tier: balanced # fast | balanced | thorough
439
- ```
268
+ ### Pipeline depth
440
269
 
441
- | Tier | Claude model | Codex model | Latency |
442
- |---|---|---|---|
443
- | `fast` | Haiku 4.5 | GPT-5.6 Luna | ~10s |
444
- | `balanced` | Sonnet 5 | GPT-5.6 Terra | ~30s |
445
- | `thorough` | Opus 4.8 | GPT-5.6 Sol | ~60s |
446
-
447
- ### Pipeline (`workflow.yml`)
270
+ The global pipeline lives in `~/.crosscheck/workflow.yml` and defaults to the full loop:
448
271
 
449
272
  ```yaml
450
273
  steps:
451
274
  - name: review
452
275
  type: review
453
- reviewer: auto # auto | claude | codex | origin
454
-
276
+ reviewer: auto # auto | claude | codex | origin
455
277
  - name: fix
456
278
  type: fix
457
279
  reviewer: origin
458
280
  when: review.verdict != 'APPROVE'
459
-
460
281
  - name: recheck
461
282
  type: recheck
462
283
  reviewer: auto
463
284
  when: fix.applied_count > 0
464
285
  ```
465
286
 
466
- ### Per-repo workflow overrides
467
-
468
- The global `workflow.yml` is the default for every repo (out of the box, the full `review → fix → recheck` loop). To run one repo at a narrower depth in the same watcher, use `crosscheck alter` — it writes a standalone override file at `~/.crosscheck/workflows/<owner>__<repo>.yml`:
287
+ To narrow a single repo without touching the global default:
469
288
 
470
289
  ```bash
471
- crosscheck alter humanbased-ai/xny-monorepo --review-only # review only
472
- crosscheck alter humanbased-ai/api --steps review,fix,recheck # full loop, explicit
473
- crosscheck alter humanbased-ai/xny-monorepo --reset # back to the global default
290
+ crosscheck alter acme/app --review-only # or --steps review,fix
474
291
  ```
475
292
 
476
- Each override file lists the review fix recheck depth only:
477
-
478
- ```yaml
479
- # ~/.crosscheck/workflows/humanbased-ai__xny-monorepo.yml
480
- steps:
481
- - review
482
- ```
293
+ That writes a standalone override at `~/.crosscheck/workflows/<owner>__<repo>.yml`, live-reloaded per PR no watcher restart.
483
294
 
484
- The override *narrows* the global workflow — it keeps each step's configured instructions and reviewer. `conflict-resolve` is orthogonal to the depth ladder: it stays enabled for any override that permits code modification (`review,fix` or `review,fix,recheck`) and is dropped only for review-only (`review`). Repos without an override file keep the complete global workflow. Resolution order: `{repo}/.crosscheck/workflow.yml` → `~/.crosscheck/workflows/<owner>__<repo>.yml` → `~/.crosscheck/workflow.yml` → built-in default.
295
+ Every option, annotated: **[crosscheck.config.example.yml](./crosscheck.config.example.yml)**.
485
296
 
486
- ### Config snapshot
487
-
488
- ```yaml
489
- # ~/.crosscheck/config.yml
490
- orgs:
491
- - your-org
297
+ ---
492
298
 
493
- routing:
494
- allowed_authors:
495
- - your-github-login
299
+ ## Running it continuously
496
300
 
497
- mode: cross-vendor # cross-vendor | single-vendor
301
+ **On your machine** a watcher for as long as your terminal is open. Webhooks arrive through a tunnel (`localhost.run` by default, zero config; `smee` if you want events queued while you're offline).
498
302
 
499
- vendors:
500
- claude:
501
- enabled: true
502
- codex:
503
- enabled: true
303
+ ```bash
304
+ crosscheck onboard && crosscheck watch
305
+ ```
504
306
 
505
- quality:
506
- tier: balanced
307
+ **On a server** — one always-on watcher for a team, with per-repo depth where it matters.
507
308
 
508
- clone_protocol: ssh # ssh (default) | https
309
+ ```bash
310
+ crosscheck onboard --team
311
+ crosscheck alter acme/legacy-service --review-only
312
+ crosscheck watch
509
313
  ```
510
314
 
511
- Full reference: [get-started.md](./get-started.md)
315
+ Deployment mode decides scope: `personal` monitors your own repos and reviews only PRs you author; `team` monitors org repos and reviews PRs from any author.
512
316
 
513
317
  ---
514
318
 
@@ -517,8 +321,8 @@ Full reference: [get-started.md](./get-started.md)
517
321
  | | Minimum |
518
322
  |---|---|
519
323
  | Node.js | 18+ |
520
- | Claude Code CLI | latest — `npm install -g @anthropic-ai/claude-code` |
521
- | Codex CLI | latest — `npm install -g @openai/codex` |
324
+ | Claude Code CLI | `npm install -g @anthropic-ai/claude-code` |
325
+ | Codex CLI | `npm install -g @openai/codex` |
522
326
  | GitHub CLI | 2.65+ — `brew install gh` |
523
327
 
524
328
  `GITHUB_TOKEN` is derived automatically from `gh auth login`. No manual export needed.
@@ -529,9 +333,13 @@ Full reference: [get-started.md](./get-started.md)
529
333
 
530
334
  | | |
531
335
  |---|---|
532
- | **[get-started.md](./get-started.md)** | Full setup guide — prerequisites, all flags, complete config reference, FAQ |
533
- | **[What 295 Agentic PRs Taught Us About Code Review](https://blog.humanbased.ai/posts/agentic-pr-quality-crosscheck/)** | Humanbased field report on agentic PR quality, review routing, and why Crosscheck exists |
534
- | **[docs/fixture-pr.md](./docs/fixture-pr.md)** | Safe public fixture PR for the first Crosscheck review |
336
+ | **[get-started.md](./get-started.md)** | Full setup guide — prerequisites, every flag, complete config reference, FAQ |
337
+ | **[docs/trust.md](./docs/trust.md)** | What leaves your machine, which GitHub permissions are needed, what Crosscheck can and cannot change, and how to try it read-only first |
338
+ | **[docs/dynamic-thoroughness.md](./docs/dynamic-thoroughness.md)** | How Crosscheck picks a model and effort per PR and the 400-PR census behind it |
339
+ | **[docs/linear-identity.md](./docs/linear-identity.md)** | Linear write-back and the attribution ladder |
340
+ | **[docs/linear-identity-contract.md](./docs/linear-identity-contract.md)** | The identity contract, as a spec for other tools |
341
+ | **[What 295 Agentic PRs Taught Us About Code Review](https://blog.humanbased.ai/posts/agentic-pr-quality-crosscheck/)** | Field report on agentic PR quality and why crosscheck exists |
342
+ | **[docs/fixture-pr.md](./docs/fixture-pr.md)** | The safe public fixture PR |
535
343
  | **[crosscheck.config.example.yml](./crosscheck.config.example.yml)** | Annotated config with every option |
536
344
  | **[CHANGELOG.md](./CHANGELOG.md)** | Release notes |
537
345
 
@@ -541,8 +349,8 @@ Full reference: [get-started.md](./get-started.md)
541
349
 
542
350
  Issues and PRs welcome at [github.com/humanbased-ai/crosscheck](https://github.com/humanbased-ai/crosscheck).
543
351
 
544
- ---
545
-
546
352
  ## License
547
353
 
548
- [MIT](./LICENSE) — Copyright (c) 2025–2026 Humanbased PTE LTD.
354
+ [MIT](./LICENSE) — Copyright (c) 2025–2026 Humanbased AI PTE LTD.
355
+
356
+ <p align="center"><em>A Humanbased project, built with crosscheck.</em></p>