@proflandrigan/shards 1.1.0

This diff represents the content of publicly available package versions that have been released to one of the supported registries. The information contained in this diff is provided for informational purposes only and reflects changes between package versions as they appear in their respective public registries.
Files changed (397) hide show
  1. package/README.md +475 -0
  2. package/package.json +37 -0
  3. package/src/agents/academic.md +276 -0
  4. package/src/agents/ai-engineer.md +377 -0
  5. package/src/agents/analytics-engineer.md +364 -0
  6. package/src/agents/applied-ml-scientist.md +410 -0
  7. package/src/agents/backend-engineer.md +255 -0
  8. package/src/agents/bi-engineer.md +333 -0
  9. package/src/agents/data-analyst.md +343 -0
  10. package/src/agents/data-engineer.md +260 -0
  11. package/src/agents/data-modeller.md +386 -0
  12. package/src/agents/data-scientist.md +366 -0
  13. package/src/agents/deep-learning-engineer.md +389 -0
  14. package/src/agents/ml-engineer.md +424 -0
  15. package/src/agents/mlops-engineer.md +339 -0
  16. package/src/agents/researcher.md +187 -0
  17. package/src/agents/specific_instructions/academic/critical_review.md +263 -0
  18. package/src/agents/specific_instructions/academic/report.md +113 -0
  19. package/src/agents/specific_instructions/ai_engineer/advise.md +162 -0
  20. package/src/agents/specific_instructions/ai_engineer/bi_engineer_handoff.md +86 -0
  21. package/src/agents/specific_instructions/ai_engineer/experiment.md +471 -0
  22. package/src/agents/specific_instructions/ai_engineer/experiment_ui_mode.md +44 -0
  23. package/src/agents/specific_instructions/ai_engineer/phases/index.md +45 -0
  24. package/src/agents/specific_instructions/ai_engineer/phases/phase-1.md +55 -0
  25. package/src/agents/specific_instructions/ai_engineer/phases/phase-2.md +86 -0
  26. package/src/agents/specific_instructions/ai_engineer/phases/phase-3.md +96 -0
  27. package/src/agents/specific_instructions/ai_engineer/phases/phase-4.md +138 -0
  28. package/src/agents/specific_instructions/ai_engineer/phases/phase-5.md +157 -0
  29. package/src/agents/specific_instructions/ai_engineer/phases/phase-6.md +196 -0
  30. package/src/agents/specific_instructions/ai_engineer/phases/phase-7.md +313 -0
  31. package/src/agents/specific_instructions/ai_engineer/phases.md +1011 -0
  32. package/src/agents/specific_instructions/ai_engineer/prompt_lab.md +161 -0
  33. package/src/agents/specific_instructions/ai_engineer/prompt_lab_ui_mode.md +28 -0
  34. package/src/agents/specific_instructions/ai_engineer/research.md +393 -0
  35. package/src/agents/specific_instructions/ai_engineer/research_ui_mode.md +66 -0
  36. package/src/agents/specific_instructions/ai_engineer/review.md +159 -0
  37. package/src/agents/specific_instructions/ai_engineer/validation_checklist.md +182 -0
  38. package/src/agents/specific_instructions/analytics_engineer/advise.md +155 -0
  39. package/src/agents/specific_instructions/analytics_engineer/bi_engineer_handoff.md +91 -0
  40. package/src/agents/specific_instructions/analytics_engineer/data_analyst_handoff.md +84 -0
  41. package/src/agents/specific_instructions/analytics_engineer/deep_phases.md +818 -0
  42. package/src/agents/specific_instructions/analytics_engineer/phases_deep/index.md +24 -0
  43. package/src/agents/specific_instructions/analytics_engineer/phases_deep/phase-1.md +77 -0
  44. package/src/agents/specific_instructions/analytics_engineer/phases_deep/phase-2.md +106 -0
  45. package/src/agents/specific_instructions/analytics_engineer/phases_deep/phase-3.md +93 -0
  46. package/src/agents/specific_instructions/analytics_engineer/phases_deep/phase-4.md +79 -0
  47. package/src/agents/specific_instructions/analytics_engineer/phases_deep/phase-5.md +61 -0
  48. package/src/agents/specific_instructions/analytics_engineer/phases_deep/phase-6.md +45 -0
  49. package/src/agents/specific_instructions/analytics_engineer/phases_deep/phase-7.md +235 -0
  50. package/src/agents/specific_instructions/analytics_engineer/phases_deep/phase-8.md +221 -0
  51. package/src/agents/specific_instructions/analytics_engineer/phases_quick/index.md +19 -0
  52. package/src/agents/specific_instructions/analytics_engineer/phases_quick/phase-1.md +47 -0
  53. package/src/agents/specific_instructions/analytics_engineer/phases_quick/phase-2.md +78 -0
  54. package/src/agents/specific_instructions/analytics_engineer/quick_phases.md +112 -0
  55. package/src/agents/specific_instructions/analytics_engineer/review.md +167 -0
  56. package/src/agents/specific_instructions/analytics_engineer/service_mode.md +369 -0
  57. package/src/agents/specific_instructions/analytics_engineer/ui_mode.md +45 -0
  58. package/src/agents/specific_instructions/analytics_engineer/update.md +162 -0
  59. package/src/agents/specific_instructions/analytics_engineer/validation_checklist.md +121 -0
  60. package/src/agents/specific_instructions/applied_ml_scientist/advise.md +143 -0
  61. package/src/agents/specific_instructions/applied_ml_scientist/phases/index.md +21 -0
  62. package/src/agents/specific_instructions/applied_ml_scientist/phases/phase-1.md +51 -0
  63. package/src/agents/specific_instructions/applied_ml_scientist/phases/phase-2.md +66 -0
  64. package/src/agents/specific_instructions/applied_ml_scientist/phases/phase-3.md +113 -0
  65. package/src/agents/specific_instructions/applied_ml_scientist/phases/phase-4.md +104 -0
  66. package/src/agents/specific_instructions/applied_ml_scientist/phases/phase-5.md +156 -0
  67. package/src/agents/specific_instructions/applied_ml_scientist/phases.md +428 -0
  68. package/src/agents/specific_instructions/applied_ml_scientist/research.md +379 -0
  69. package/src/agents/specific_instructions/applied_ml_scientist/review.md +142 -0
  70. package/src/agents/specific_instructions/applied_ml_scientist/validation_checklist.md +136 -0
  71. package/src/agents/specific_instructions/backend_engineer/clean.md +149 -0
  72. package/src/agents/specific_instructions/backend_engineer/review.md +91 -0
  73. package/src/agents/specific_instructions/backend_engineer/review_checklist.md +54 -0
  74. package/src/agents/specific_instructions/backend_engineer/service_mode.md +67 -0
  75. package/src/agents/specific_instructions/bi_engineer/advise.md +137 -0
  76. package/src/agents/specific_instructions/bi_engineer/data_analyst_handoff.md +77 -0
  77. package/src/agents/specific_instructions/bi_engineer/incoming_handoff.md +45 -0
  78. package/src/agents/specific_instructions/bi_engineer/phases/index.md +20 -0
  79. package/src/agents/specific_instructions/bi_engineer/phases/phase-1.md +164 -0
  80. package/src/agents/specific_instructions/bi_engineer/phases/phase-2.md +92 -0
  81. package/src/agents/specific_instructions/bi_engineer/phases/phase-3.md +121 -0
  82. package/src/agents/specific_instructions/bi_engineer/phases/phase-4.md +106 -0
  83. package/src/agents/specific_instructions/bi_engineer/phases.md +451 -0
  84. package/src/agents/specific_instructions/bi_engineer/review.md +166 -0
  85. package/src/agents/specific_instructions/bi_engineer/update.md +147 -0
  86. package/src/agents/specific_instructions/bi_engineer/validation_checklist.md +124 -0
  87. package/src/agents/specific_instructions/data_analyst/advise.md +138 -0
  88. package/src/agents/specific_instructions/data_analyst/explain.md +221 -0
  89. package/src/agents/specific_instructions/data_analyst/incoming_handoff.md +40 -0
  90. package/src/agents/specific_instructions/data_analyst/phases/index.md +20 -0
  91. package/src/agents/specific_instructions/data_analyst/phases/phase-1.md +159 -0
  92. package/src/agents/specific_instructions/data_analyst/phases/phase-2.md +112 -0
  93. package/src/agents/specific_instructions/data_analyst/phases/phase-3.md +265 -0
  94. package/src/agents/specific_instructions/data_analyst/phases/phase-4.md +100 -0
  95. package/src/agents/specific_instructions/data_analyst/phases.md +501 -0
  96. package/src/agents/specific_instructions/data_analyst/review.md +138 -0
  97. package/src/agents/specific_instructions/data_analyst/ui_mode.md +26 -0
  98. package/src/agents/specific_instructions/data_analyst/update.md +144 -0
  99. package/src/agents/specific_instructions/data_analyst/validation_checklist.md +95 -0
  100. package/src/agents/specific_instructions/data_engineer/advise.md +137 -0
  101. package/src/agents/specific_instructions/data_engineer/phases.md +466 -0
  102. package/src/agents/specific_instructions/data_engineer/phases_deep/index.md +23 -0
  103. package/src/agents/specific_instructions/data_engineer/phases_deep/phase-1.md +49 -0
  104. package/src/agents/specific_instructions/data_engineer/phases_deep/phase-2.md +93 -0
  105. package/src/agents/specific_instructions/data_engineer/phases_deep/phase-3.md +55 -0
  106. package/src/agents/specific_instructions/data_engineer/phases_deep/phase-4.md +48 -0
  107. package/src/agents/specific_instructions/data_engineer/phases_deep/phase-5.md +40 -0
  108. package/src/agents/specific_instructions/data_engineer/phases_deep/phase-6.md +102 -0
  109. package/src/agents/specific_instructions/data_engineer/phases_deep/phase-7.md +87 -0
  110. package/src/agents/specific_instructions/data_engineer/phases_quick/index.md +19 -0
  111. package/src/agents/specific_instructions/data_engineer/phases_quick/phase-1.md +45 -0
  112. package/src/agents/specific_instructions/data_engineer/phases_quick/phase-2.md +54 -0
  113. package/src/agents/specific_instructions/data_engineer/review.md +135 -0
  114. package/src/agents/specific_instructions/data_engineer/validation_checklist.md +136 -0
  115. package/src/agents/specific_instructions/data_modeller/advise.md +137 -0
  116. package/src/agents/specific_instructions/data_modeller/phases.md +581 -0
  117. package/src/agents/specific_instructions/data_modeller/phases_deep/index.md +23 -0
  118. package/src/agents/specific_instructions/data_modeller/phases_deep/phase-1.md +52 -0
  119. package/src/agents/specific_instructions/data_modeller/phases_deep/phase-2.md +113 -0
  120. package/src/agents/specific_instructions/data_modeller/phases_deep/phase-3.md +47 -0
  121. package/src/agents/specific_instructions/data_modeller/phases_deep/phase-4.md +51 -0
  122. package/src/agents/specific_instructions/data_modeller/phases_deep/phase-5.md +45 -0
  123. package/src/agents/specific_instructions/data_modeller/phases_deep/phase-6.md +105 -0
  124. package/src/agents/specific_instructions/data_modeller/phases_deep/phase-7.md +136 -0
  125. package/src/agents/specific_instructions/data_modeller/phases_quick/index.md +19 -0
  126. package/src/agents/specific_instructions/data_modeller/phases_quick/phase-1.md +47 -0
  127. package/src/agents/specific_instructions/data_modeller/phases_quick/phase-2.md +65 -0
  128. package/src/agents/specific_instructions/data_modeller/review.md +141 -0
  129. package/src/agents/specific_instructions/data_modeller/service_mode.md +218 -0
  130. package/src/agents/specific_instructions/data_modeller/validation_checklist.md +125 -0
  131. package/src/agents/specific_instructions/data_scientist/advise.md +158 -0
  132. package/src/agents/specific_instructions/data_scientist/bi_engineer_handoff.md +63 -0
  133. package/src/agents/specific_instructions/data_scientist/experiment.md +482 -0
  134. package/src/agents/specific_instructions/data_scientist/experiment_ui_mode.md +44 -0
  135. package/src/agents/specific_instructions/data_scientist/explain.md +247 -0
  136. package/src/agents/specific_instructions/data_scientist/greenfield_data.md +35 -0
  137. package/src/agents/specific_instructions/data_scientist/ml_engineer_handoff.md +52 -0
  138. package/src/agents/specific_instructions/data_scientist/notebook_walkthrough.md +76 -0
  139. package/src/agents/specific_instructions/data_scientist/phases/index.md +24 -0
  140. package/src/agents/specific_instructions/data_scientist/phases/phase-1.md +45 -0
  141. package/src/agents/specific_instructions/data_scientist/phases/phase-2.md +67 -0
  142. package/src/agents/specific_instructions/data_scientist/phases/phase-3.md +89 -0
  143. package/src/agents/specific_instructions/data_scientist/phases/phase-4.md +143 -0
  144. package/src/agents/specific_instructions/data_scientist/phases/phase-5.md +71 -0
  145. package/src/agents/specific_instructions/data_scientist/phases/phase-6.md +239 -0
  146. package/src/agents/specific_instructions/data_scientist/phases/phase-7.md +207 -0
  147. package/src/agents/specific_instructions/data_scientist/phases.md +651 -0
  148. package/src/agents/specific_instructions/data_scientist/research.md +345 -0
  149. package/src/agents/specific_instructions/data_scientist/research_ui_mode.md +52 -0
  150. package/src/agents/specific_instructions/data_scientist/review.md +136 -0
  151. package/src/agents/specific_instructions/data_scientist/service_mode.md +247 -0
  152. package/src/agents/specific_instructions/data_scientist/validation_checklist.md +183 -0
  153. package/src/agents/specific_instructions/deep_learning_engineer/advise.md +145 -0
  154. package/src/agents/specific_instructions/deep_learning_engineer/phases/index.md +21 -0
  155. package/src/agents/specific_instructions/deep_learning_engineer/phases/phase-1.md +74 -0
  156. package/src/agents/specific_instructions/deep_learning_engineer/phases/phase-2.md +98 -0
  157. package/src/agents/specific_instructions/deep_learning_engineer/phases/phase-3.md +76 -0
  158. package/src/agents/specific_instructions/deep_learning_engineer/phases/phase-4.md +128 -0
  159. package/src/agents/specific_instructions/deep_learning_engineer/phases/phase-5.md +292 -0
  160. package/src/agents/specific_instructions/deep_learning_engineer/phases.md +567 -0
  161. package/src/agents/specific_instructions/deep_learning_engineer/research.md +389 -0
  162. package/src/agents/specific_instructions/deep_learning_engineer/review.md +155 -0
  163. package/src/agents/specific_instructions/deep_learning_engineer/validation_checklist.md +147 -0
  164. package/src/agents/specific_instructions/ml_engineer/advise.md +174 -0
  165. package/src/agents/specific_instructions/ml_engineer/bi_engineer_handoff.md +71 -0
  166. package/src/agents/specific_instructions/ml_engineer/experiment.md +474 -0
  167. package/src/agents/specific_instructions/ml_engineer/experiment_ui_mode.md +44 -0
  168. package/src/agents/specific_instructions/ml_engineer/notebook_walkthrough.md +75 -0
  169. package/src/agents/specific_instructions/ml_engineer/phases/index.md +25 -0
  170. package/src/agents/specific_instructions/ml_engineer/phases/phase-1.md +49 -0
  171. package/src/agents/specific_instructions/ml_engineer/phases/phase-2.md +75 -0
  172. package/src/agents/specific_instructions/ml_engineer/phases/phase-3.md +124 -0
  173. package/src/agents/specific_instructions/ml_engineer/phases/phase-4.md +279 -0
  174. package/src/agents/specific_instructions/ml_engineer/phases/phase-5.md +160 -0
  175. package/src/agents/specific_instructions/ml_engineer/phases/phase-6-5.md +170 -0
  176. package/src/agents/specific_instructions/ml_engineer/phases/phase-6.md +295 -0
  177. package/src/agents/specific_instructions/ml_engineer/phases/phase-7.md +337 -0
  178. package/src/agents/specific_instructions/ml_engineer/phases.md +1068 -0
  179. package/src/agents/specific_instructions/ml_engineer/research.md +437 -0
  180. package/src/agents/specific_instructions/ml_engineer/research_ui_mode.md +71 -0
  181. package/src/agents/specific_instructions/ml_engineer/review.md +187 -0
  182. package/src/agents/specific_instructions/ml_engineer/service_mode.md +273 -0
  183. package/src/agents/specific_instructions/ml_engineer/validation_checklist.md +185 -0
  184. package/src/agents/specific_instructions/mlops_engineer/advise.md +139 -0
  185. package/src/agents/specific_instructions/mlops_engineer/phases/index.md +23 -0
  186. package/src/agents/specific_instructions/mlops_engineer/phases/phase-1.md +52 -0
  187. package/src/agents/specific_instructions/mlops_engineer/phases/phase-2.md +86 -0
  188. package/src/agents/specific_instructions/mlops_engineer/phases/phase-3.md +105 -0
  189. package/src/agents/specific_instructions/mlops_engineer/phases/phase-4.md +128 -0
  190. package/src/agents/specific_instructions/mlops_engineer/phases/phase-5.md +106 -0
  191. package/src/agents/specific_instructions/mlops_engineer/phases/phase-6.md +128 -0
  192. package/src/agents/specific_instructions/mlops_engineer/phases/phase-7.md +144 -0
  193. package/src/agents/specific_instructions/mlops_engineer/phases.md +671 -0
  194. package/src/agents/specific_instructions/mlops_engineer/review.md +164 -0
  195. package/src/agents/specific_instructions/mlops_engineer/service_mode.md +81 -0
  196. package/src/agents/specific_instructions/mlops_engineer/validation_checklist.md +151 -0
  197. package/src/agents/specific_instructions/researcher/critical_review.md +292 -0
  198. package/src/agents/specific_instructions/researcher/review_checklist.md +67 -0
  199. package/src/agents/specific_instructions/researcher/service_mode.md +224 -0
  200. package/src/agents/specific_instructions/shared/auto_verify_mode.md +141 -0
  201. package/src/agents/specific_instructions/shared/autonomous_research.md +1289 -0
  202. package/src/agents/specific_instructions/shared/behavioral_rules.md +36 -0
  203. package/src/agents/specific_instructions/shared/diverge_protocol.md +387 -0
  204. package/src/agents/specific_instructions/shared/engineering_guidelines.md +136 -0
  205. package/src/agents/specific_instructions/shared/experiment_versioning.md +184 -0
  206. package/src/agents/specific_instructions/shared/goal_mode.md +187 -0
  207. package/src/agents/specific_instructions/shared/incremental_testing.md +139 -0
  208. package/src/agents/specific_instructions/shared/intent_discovery.md +223 -0
  209. package/src/agents/specific_instructions/shared/join_path_protocol.md +168 -0
  210. package/src/agents/specific_instructions/shared/knowledge_checkpoint.md +83 -0
  211. package/src/agents/specific_instructions/shared/knowledge_harvest.md +220 -0
  212. package/src/agents/specific_instructions/shared/knowledge_retrieval.md +100 -0
  213. package/src/agents/specific_instructions/shared/notebook_walkthrough_protocol.md +367 -0
  214. package/src/agents/specific_instructions/shared/reviewer_verdict_protocol.md +74 -0
  215. package/src/agents/specific_instructions/shared/swarm_protocol.md +97 -0
  216. package/src/agents/specific_instructions/shared/validation_protocol.md +139 -0
  217. package/src/agents/specific_instructions/syn/arbiter.md +140 -0
  218. package/src/agents/specific_instructions/syn/brainstorm.md +550 -0
  219. package/src/agents/specific_instructions/syn/code_review.md +232 -0
  220. package/src/agents/specific_instructions/syn/diff.md +239 -0
  221. package/src/agents/specific_instructions/syn/final_review.md +65 -0
  222. package/src/agents/specific_instructions/syn/fixer.md +240 -0
  223. package/src/agents/specific_instructions/syn/free_form.md +130 -0
  224. package/src/agents/specific_instructions/syn/knowledge.md +468 -0
  225. package/src/agents/specific_instructions/syn/notebook_walkthrough.md +78 -0
  226. package/src/agents/specific_instructions/syn/panel_review.md +634 -0
  227. package/src/agents/specific_instructions/syn/pm.md +453 -0
  228. package/src/agents/specific_instructions/syn/pr_review.md +255 -0
  229. package/src/agents/specific_instructions/syn/slides.md +417 -0
  230. package/src/agents/syn.md +729 -0
  231. package/src/commands/academic.md +41 -0
  232. package/src/commands/ai-engineer.md +45 -0
  233. package/src/commands/analytics-engineer.md +48 -0
  234. package/src/commands/applied-ml-scientist.md +45 -0
  235. package/src/commands/backend-engineer.md +35 -0
  236. package/src/commands/bi-engineer.md +40 -0
  237. package/src/commands/brainstorm.md +24 -0
  238. package/src/commands/data-analyst.md +38 -0
  239. package/src/commands/data-engineer.md +37 -0
  240. package/src/commands/data-modeller.md +38 -0
  241. package/src/commands/data-scientist.md +38 -0
  242. package/src/commands/deep-learning-engineer.md +47 -0
  243. package/src/commands/end.md +49 -0
  244. package/src/commands/knowledge.md +24 -0
  245. package/src/commands/ml-engineer.md +42 -0
  246. package/src/commands/mlops-engineer.md +47 -0
  247. package/src/commands/notebook-walkthrough.md +58 -0
  248. package/src/commands/researcher.md +40 -0
  249. package/src/commands/resume.md +57 -0
  250. package/src/commands/review-pr.md +26 -0
  251. package/src/commands/shards-guide.md +41 -0
  252. package/src/commands/shards-ui.md +32 -0
  253. package/src/commands/shards.md +41 -0
  254. package/src/docs/01-getting-started/concepts.md +109 -0
  255. package/src/docs/01-getting-started/first-session.md +79 -0
  256. package/src/docs/01-getting-started/install.md +61 -0
  257. package/src/docs/02-agents/academic.md +71 -0
  258. package/src/docs/02-agents/ai-engineer.md +78 -0
  259. package/src/docs/02-agents/analytics-engineer.md +58 -0
  260. package/src/docs/02-agents/applied-ml-scientist.md +59 -0
  261. package/src/docs/02-agents/backend-engineer.md +58 -0
  262. package/src/docs/02-agents/bi-engineer.md +65 -0
  263. package/src/docs/02-agents/data-analyst.md +67 -0
  264. package/src/docs/02-agents/data-engineer.md +57 -0
  265. package/src/docs/02-agents/data-modeller.md +51 -0
  266. package/src/docs/02-agents/data-scientist.md +78 -0
  267. package/src/docs/02-agents/deep-learning-engineer.md +64 -0
  268. package/src/docs/02-agents/ml-engineer.md +80 -0
  269. package/src/docs/02-agents/mlops-engineer.md +59 -0
  270. package/src/docs/02-agents/overview.md +62 -0
  271. package/src/docs/02-agents/researcher.md +73 -0
  272. package/src/docs/02-agents/syn.md +88 -0
  273. package/src/docs/03-protocols/auto-verify.md +82 -0
  274. package/src/docs/03-protocols/autonomous-research.md +59 -0
  275. package/src/docs/03-protocols/behavioral-rules.md +35 -0
  276. package/src/docs/03-protocols/diverge.md +50 -0
  277. package/src/docs/03-protocols/engineering-guidelines.md +56 -0
  278. package/src/docs/03-protocols/experiment-versioning.md +38 -0
  279. package/src/docs/03-protocols/gate-pattern.md +65 -0
  280. package/src/docs/03-protocols/incremental-testing.md +68 -0
  281. package/src/docs/03-protocols/join-path.md +46 -0
  282. package/src/docs/03-protocols/knowledge-ledger.md +70 -0
  283. package/src/docs/03-protocols/reviewer-verdicts.md +39 -0
  284. package/src/docs/03-protocols/swarm.md +40 -0
  285. package/src/docs/03-protocols/validation.md +174 -0
  286. package/src/docs/04-ui/activity-bar.md +70 -0
  287. package/src/docs/04-ui/chat-pane.md +80 -0
  288. package/src/docs/04-ui/code-intel.md +62 -0
  289. package/src/docs/04-ui/file-editing.md +61 -0
  290. package/src/docs/04-ui/git.md +54 -0
  291. package/src/docs/04-ui/keybindings.md +79 -0
  292. package/src/docs/04-ui/knowledge-map.md +76 -0
  293. package/src/docs/04-ui/overview.md +93 -0
  294. package/src/docs/04-ui/panels.md +49 -0
  295. package/src/docs/04-ui/pinboard-selection.md +66 -0
  296. package/src/docs/04-ui/quick-open-palette.md +56 -0
  297. package/src/docs/04-ui/sessions.md +81 -0
  298. package/src/docs/04-ui/settings-permissions.md +56 -0
  299. package/src/docs/05-commands/reference.md +59 -0
  300. package/src/docs/06-outputs/directory-map.md +116 -0
  301. package/src/docs/07-workflows/ai-eval-first.md +57 -0
  302. package/src/docs/07-workflows/deep-study-to-production.md +76 -0
  303. package/src/docs/07-workflows/diverge-exploration.md +77 -0
  304. package/src/docs/07-workflows/quick-analysis.md +45 -0
  305. package/src/docs/08-integrations/claude-code-auto-mode.md +191 -0
  306. package/src/docs/08-integrations/google-slides.md +175 -0
  307. package/src/docs/README.md +30 -0
  308. package/src/docs/manifest.json +108 -0
  309. package/src/templates/analysis-template.md +20 -0
  310. package/src/templates/branch-report.md +46 -0
  311. package/src/templates/diff-report.md +88 -0
  312. package/src/templates/knowledge-index.md +7 -0
  313. package/src/templates/model-card-schema.json +186 -0
  314. package/src/templates/model-card-schema.md +88 -0
  315. package/src/templates/model-card.md +124 -0
  316. package/src/templates/project-plan.md +47 -0
  317. package/src/templates/project-specs.md +81 -0
  318. package/src/templates/report-template.md +43 -0
  319. package/src/templates/study-template.md +25 -0
  320. package/src/ui/cc-readonly.js +181 -0
  321. package/src/ui/chat-session.js +466 -0
  322. package/src/ui/css/base.css +136 -0
  323. package/src/ui/css/brainstorm.css +525 -0
  324. package/src/ui/css/chat.css +1405 -0
  325. package/src/ui/css/editor.css +546 -0
  326. package/src/ui/css/eval-dashboard.css +157 -0
  327. package/src/ui/css/experiment.css +237 -0
  328. package/src/ui/css/guide.css +186 -0
  329. package/src/ui/css/knowledge-map.css +383 -0
  330. package/src/ui/css/layout.css +431 -0
  331. package/src/ui/css/model-card.css +161 -0
  332. package/src/ui/css/notebook-walkthrough.css +271 -0
  333. package/src/ui/css/pr-review.css +403 -0
  334. package/src/ui/css/prompt-lab.css +325 -0
  335. package/src/ui/css/sessions.css +258 -0
  336. package/src/ui/css/sidebar.css +661 -0
  337. package/src/ui/css/terminal.css +113 -0
  338. package/src/ui/css/theme-light.css +542 -0
  339. package/src/ui/index.html +389 -0
  340. package/src/ui/js/agents.js +32 -0
  341. package/src/ui/js/bookmarks.js +230 -0
  342. package/src/ui/js/chat.js +1776 -0
  343. package/src/ui/js/code-intel.js +328 -0
  344. package/src/ui/js/command-palette.js +142 -0
  345. package/src/ui/js/events.js +591 -0
  346. package/src/ui/js/explorer.js +317 -0
  347. package/src/ui/js/file-view.js +477 -0
  348. package/src/ui/js/git.js +536 -0
  349. package/src/ui/js/guide.js +198 -0
  350. package/src/ui/js/hud.js +75 -0
  351. package/src/ui/js/init.js +351 -0
  352. package/src/ui/js/knowledge-map.js +906 -0
  353. package/src/ui/js/markdown.js +114 -0
  354. package/src/ui/js/monaco.js +164 -0
  355. package/src/ui/js/notebook-walkthrough.js +272 -0
  356. package/src/ui/js/notebook.js +448 -0
  357. package/src/ui/js/panels.js +2681 -0
  358. package/src/ui/js/pinboard.js +186 -0
  359. package/src/ui/js/quick-open.js +164 -0
  360. package/src/ui/js/selection-context.js +131 -0
  361. package/src/ui/js/sessions.js +256 -0
  362. package/src/ui/js/settings.js +476 -0
  363. package/src/ui/js/split-view.js +82 -0
  364. package/src/ui/js/state.js +343 -0
  365. package/src/ui/js/table.js +161 -0
  366. package/src/ui/js/tabs.js +284 -0
  367. package/src/ui/js/tabular.js +125 -0
  368. package/src/ui/js/terminal.js +354 -0
  369. package/src/ui/js/timeline.js +137 -0
  370. package/src/ui/js/utils.js +293 -0
  371. package/src/ui/notebook-kernel.py +790 -0
  372. package/src/ui/open-browser.js +55 -0
  373. package/src/ui/permission-pattern.js +42 -0
  374. package/src/ui/relay.js +513 -0
  375. package/src/ui/server.js +3072 -0
  376. package/src/ui/session-index.js +225 -0
  377. package/src/ui/shards_icon.png +0 -0
  378. package/src/ui/spawn-server.js +41 -0
  379. package/src/ui/symbol-index.js +813 -0
  380. package/src/ui/ui-push.js +177 -0
  381. package/tools/gate-hook/VALIDATION_SPEC.md +273 -0
  382. package/tools/gate-hook/__tests__/auto-verify.test.js +343 -0
  383. package/tools/gate-hook/auto-allowlist.js +179 -0
  384. package/tools/gate-hook/auto-state.js +68 -0
  385. package/tools/gate-hook/classify.js +21 -0
  386. package/tools/gate-hook/log.js +57 -0
  387. package/tools/gate-hook/parser.js +205 -0
  388. package/tools/gate-hook/sql-guard.js +230 -0
  389. package/tools/gate-hook/state.js +170 -0
  390. package/tools/gate-hook/sweep.js +139 -0
  391. package/tools/gate-hook/transcript.js +45 -0
  392. package/tools/gate-hook/validation.js +321 -0
  393. package/tools/gate-hook.js +475 -0
  394. package/tools/install.js +914 -0
  395. package/tools/shards-gates.js +311 -0
  396. package/tools/shards-sessions.js +261 -0
  397. package/tools/shards-ui.js +377 -0
@@ -0,0 +1,104 @@
1
+ > **Previous:** phase-3.md confirmed
2
+ > **Next:** phase-5.md (read only after this phase's gate is confirmed)
3
+
4
+ ---
5
+
6
+ ## Create Mode — Phase 4: Execute (Gated)
7
+
8
+ Goal: Build the prototype.
9
+
10
+ **Context checkpoint:** Before building, prompt the user:
11
+
12
+ "Planning's locked — good moment to run `/compact` or `/clear` before we start
13
+ executing. I'll be working from project-specs.md from here. Say the word when
14
+ you're ready."
15
+
16
+ Wait for any signal from the user before beginning build steps.
17
+
18
+ **Knowledge re-check:** Follow `.claude/agents/specific_instructions/shared/knowledge_checkpoint.md` before building.
19
+
20
+ ### Incremental testing — checkpoint gates between components
21
+
22
+ Follow `.claude/agents/specific_instructions/shared/incremental_testing.md` during this build. Each component below is a checkpoint seam — after you write and execute a component, emit a `kind=checkpoint` gate fence (template below) and wait for user confirmation before starting the next component. Do not leave run-all until the end: test each component in isolation as you build it.
23
+
24
+ Checkpoint gate fence — emit exactly this shape. Both `::GATE::` and `::ENDGATE::` fences are required, as are all three attributes (`id`, `phase`, `kind`). No prose outside the fence.
25
+
26
+ ```
27
+ ::GATE:: id=<agent-name>-phase-<N>-checkpoint-<component> phase=<N> kind=checkpoint
28
+ Component: <human-readable name>
29
+ Test command: <exact command you ran>
30
+ Evidence:
31
+ - <measured fact 1, e.g. "df.shape = (48211, 47)">
32
+ - <measured fact 2, e.g. "null rate on join key = 0.00%">
33
+ - <measured fact 3, e.g. "sample head matches expected schema">
34
+ Status: PASS | FAIL — <one-line summary>
35
+ Next: <what you'll build after this is confirmed>
36
+ Stop here — await explicit confirmation before writing the next component.
37
+ ::ENDGATE::
38
+ ```
39
+
40
+ Expected checkpoint gate IDs for this phase (emit in order as you build):
41
+
42
+ - `applied-ml-scientist-phase-4-checkpoint-data` — data / synthetic-data cell produces expected shape; a sample inspection confirms structure.
43
+ - `applied-ml-scientist-phase-4-checkpoint-components` — each framework component forward-passes on dummy input with correct output shape and dtype.
44
+ - `applied-ml-scientist-phase-4-checkpoint-smoke-train` — training loop runs for 10-50 steps on a tiny batch; loss decreases (not flat, not NaN).
45
+ - `applied-ml-scientist-phase-4-checkpoint-full-train` — full training completes; loss curve plotted; convergence direction matches prediction.
46
+ - `applied-ml-scientist-phase-4-checkpoint-eval` — evaluation vs baselines runs; metric table coherent; ablation (if feasible) logged.
47
+
48
+ The hook blocks all non-read tools while a checkpoint is open. If a checkpoint fails, diagnose and re-emit with updated evidence before advancing. Use the fence body format shown above (Component / Test command / Evidence / Status / Next).
49
+
50
+ **Create `research/<project_name>/notebooks/framework_prototype.ipynb`:**
51
+
52
+ Structure the notebook with these sections:
53
+ 1. **Setup** — imports, configuration, device setup, seed setting
54
+ 2. **Data** — data loading (real) or synthetic data generation; EDA/visualization
55
+ of a sample to confirm structure
56
+ 3. **Framework Implementation** — implement each core component cell by cell,
57
+ with markdown explaining each component's role and design choices
58
+ 4. **Training Loop** — full training loop with logging; run for enough steps to
59
+ verify gradient flow and loss convergence direction
60
+ 5. **Evaluation** — run against baselines; produce metric table; ablation runs
61
+ if feasible in the prototype
62
+ 6. **Visualization** — loss curves, learned representations (t-SNE/UMAP if
63
+ applicable), attention maps, or whatever is diagnostic for this architecture
64
+
65
+ **Create `research/<project_name>/src/` module files:**
66
+
67
+ Extract reusable components from the notebook into proper Python modules.
68
+ Each module should be importable and have clean interfaces. Prefer explicit
69
+ over clever.
70
+
71
+ **Create `research/<project_name>/requirements.txt`**
72
+
73
+ ### Document Phase 4
74
+
75
+ Append to `project-specs.md`:
76
+
77
+ ```markdown
78
+ ## Phase 4: Build Log
79
+
80
+ - **Notebook:** `notebooks/framework_prototype.ipynb`
81
+ - **Modules created:** <list of src/ files>
82
+ - **Training run summary:**
83
+ - Steps / epochs: <N>
84
+ - Hardware: <GPU/CPU>
85
+ - Training loss trajectory: <converged | diverged | oscillating — describe>
86
+ - Validation metric: <value>
87
+ - **Baseline comparison:**
88
+ | Method | Metric | Notes |
89
+ |--------|--------|-------|
90
+ | <baseline> | <value> | — |
91
+ | **Ours** | <value> | — |
92
+ - **Known limitations:** <what the prototype doesn't handle yet>
93
+ - **Ablation results (if run):** <summary>
94
+ ```
95
+
96
+ ::GATE:: id=applied-ml-scientist-phase-4 phase=4 kind=phase validates=applied_ml_scientist
97
+ Read this section back to the user. Stop here — do not begin the next phase or output any further content. Wait for the user to explicitly confirm before proceeding. Do not interpret silence or partial agreement as confirmation.
98
+ ::ENDGATE::
99
+
100
+ ---
101
+
102
+ ## When this gate is confirmed
103
+
104
+ Read `.claude/agents/specific_instructions/applied_ml_scientist/phases/phase-5.md` in full and follow its instructions starting from Phase 5. Do not pre-read further phase files.
@@ -0,0 +1,156 @@
1
+ > **Previous:** phase-4.md confirmed
2
+ > **Next:** This is the final phase — follow the Syn sign-off instructions in this phase to close the project.
3
+
4
+ ---
5
+
6
+ ## Create Mode — Phase 5: Review & Handoff (Gated)
7
+
8
+ Goal: Final review, report, and sign-off.
9
+
10
+ **If the framework involves deep learning components** — custom neural
11
+ architectures, specialized training objectives for neural models, or
12
+ implementation of DL-based novel methods — consult the Deep Learning Engineer
13
+ for implementation grounding before the Syn review:
14
+
15
+ Tell the user: "This framework has deep learning implementation requirements —
16
+ I'm asking the Deep Learning Engineer shard to review implementation fidelity,
17
+ tensor operations, and numerical stability before we close..."
18
+
19
+ ```
20
+ Task(
21
+ subagent_type="deep-learning-engineer",
22
+ description="DL implementation review for novel ML framework: <project name>",
23
+ prompt="I am the Applied ML Scientist shard. I have designed a novel ML
24
+ framework with deep learning components and need an implementation review
25
+ before final sign-off.
26
+
27
+ Project: <project name>
28
+ Directory: research/<project_name>/
29
+ Specs: research/<project_name>/project-specs.md
30
+
31
+ Framework summary:
32
+ - Novel contribution: <core hypothesis from Phase 1>
33
+ - Core DL components: <custom architectures or mechanisms from Phase 2>
34
+ - Training objective: <loss function formula from Phase 2>
35
+ - Framework: <PyTorch | JAX from Phase 3>
36
+ - Data modality: <from Phase 0>
37
+ - Scale: <N examples, sequence/spatial dims>
38
+
39
+ Please review for implementation fidelity:
40
+ 1. Are the custom differentiable operations correctly implementable in the
41
+ chosen framework without approximation errors?
42
+ 2. Are there numerical instability risks in the proposed architecture or
43
+ loss function (softmax overflow, vanishing gradients, BatchNorm at small
44
+ batch sizes, etc.)?
45
+ 3. Are the tensor shapes and operations consistent through the full forward
46
+ pass as described?
47
+ 4. What is the memory and compute cost estimate, and does it fit the stated
48
+ hardware constraints?
49
+ 5. Are there implementation-level gaps between the theoretical design and
50
+ what is practically achievable with current tooling?
51
+
52
+ Please read project-specs.md for full context."
53
+ )
54
+ ```
55
+
56
+ Address any blocking implementation concerns raised before proceeding to Syn.
57
+
58
+ **Backend Engineer code review (Python artifacts):**
59
+
60
+ Glob the project directory (`research/<project_name>/`) for `.py` and `.ipynb` files.
61
+ If any are found:
62
+
63
+ Tell the user: "Before Syn signs off, the Backend Engineer is reviewing the
64
+ Python artifacts. Code quality is not optional."
65
+
66
+ ```
67
+ Task(
68
+ subagent_type="backend-engineer",
69
+ description="Python code review for [project_name]",
70
+ prompt="You are in SERVICE MODE. Review the following Python files in the
71
+ project at research/[project_name]/. Read project-specs.md first for context.
72
+ Files to review: [list of .py and .ipynb files found]"
73
+ )
74
+ ```
75
+
76
+ Append the Backend Engineer's review to project-specs.md. If no Python files are
77
+ found, skip this step.
78
+
79
+ **Consult Syn for final sign-off:**
80
+
81
+ Tell the user: "I'm asking Syn to review the framework design and results
82
+ before we close..."
83
+
84
+ ```
85
+ Task(
86
+ subagent_type="syn",
87
+ description="Final review of novel ML framework: <project name>",
88
+ prompt="I am the Applied ML Scientist shard. I have completed a novel ML
89
+ framework research project. Please review and provide APPROVED / NEEDS
90
+ REVISION / BLOCKED.
91
+
92
+ Project: <project name>
93
+ Directory: research/<project_name>/
94
+ Specs: research/<project_name>/project-specs.md
95
+
96
+ Summary:
97
+ - Problem: <one sentence from Phase 0>
98
+ - Novel contribution: <core hypothesis from Phase 1>
99
+ - Architecture: <key components from Phase 2>
100
+ - Results: <baseline comparison summary from Phase 4>
101
+ - Known limitations: <from Phase 4>
102
+
103
+ Please read project-specs.md for full context."
104
+ )
105
+ ```
106
+
107
+ **Create `research/<project_name>/report.md`:**
108
+
109
+ ```markdown
110
+ # <Project Name> — Research Report
111
+
112
+ ## Executive Summary
113
+ <2-3 sentences: what was built, why it's novel, and what the results show>
114
+
115
+ ## Novel Contribution
116
+ <What specifically is new here, stated precisely at the component or
117
+ objective level — not "we achieve better performance" but "we introduce X
118
+ mechanism which encodes Y inductive bias, enabling Z capability">
119
+
120
+ ## Results
121
+ <Metric table comparing to baselines>
122
+ <Key training dynamics observations>
123
+ <Ablation results if available>
124
+
125
+ ## Limitations
126
+ <What the prototype doesn't handle, what remains unvalidated,
127
+ scale constraints, data quality assumptions>
128
+
129
+ ## Code Review
130
+ **Backend Engineer verdict:** <Clean | Minor Issues | Refactor Required | Blocked | N/A — no Python artifacts>
131
+ <Summary of code review findings, or "No Python artifacts found.">
132
+
133
+ ## Next Steps
134
+ <Ordered by expected impact:>
135
+ 1. <experiment or engineering step>
136
+ 2. <experiment or engineering step>
137
+ 3. <experiment or engineering step>
138
+
139
+ ## Knowledge Harvested
140
+ - <title> → .shards/knowledge/<type>/<filename>.md
141
+ - Or: None — project did not produce reusable knowledge
142
+ ```
143
+
144
+ **Knowledge harvest.** Before closing, extract reusable knowledge from this project.
145
+ Read `.claude/agents/specific_instructions/shared/knowledge_harvest.md` and follow
146
+ the protocol. Present candidates to the user for confirmation before writing.
147
+
148
+ ::GATE:: id=applied-ml-scientist-phase-5 phase=5 kind=final
149
+ Read Phase 5 summary to the user. Stop here — wait for the user to explicitly confirm the project is closed before wrapping up.
150
+ ::ENDGATE::
151
+
152
+ ---
153
+
154
+ ## When this gate is confirmed
155
+
156
+ This is the final phase. Once Syn returns APPROVED sign-off, the project is complete. Do not read further files.
@@ -0,0 +1,428 @@
1
+ # Applied ML Scientist — Phased Workflow (Create Mode)
2
+
3
+ Phases 1 through 5 for the Applied ML Scientist Create Mode.
4
+ Phase 0 (Problem Framing) is already complete.
5
+ Follow every phase, gate, and documentation rule below.
6
+
7
+ ---
8
+
9
+ ## Create Mode — Phase 1: Research Landscape (Gated)
10
+
11
+ Goal: Map the design space. Understand what exists before defining what's novel.
12
+
13
+ 1. Identify the 3-5 most relevant methods or papers from the literature for
14
+ this problem
15
+ 2. For each: what does it do well, and where specifically does it break down?
16
+ 3. Identify the gap the novel framework will fill — what property does none of
17
+ the existing methods have?
18
+ 4. Articulate the core hypothesis: *what structural insight makes the new
19
+ approach work where others don't?*
20
+
21
+ Present findings conversationally before documenting. Ask the user if any of
22
+ the surveyed methods are ones they've already evaluated and ruled out.
23
+
24
+ ### Document Phase 1
25
+
26
+ Append to `project-specs.md`:
27
+
28
+ ```markdown
29
+ ## Phase 1: Research Landscape
30
+
31
+ ### Relevant Prior Work
32
+ | Method / Paper | Core Idea | Strengths | Failure Modes Relevant to Our Problem |
33
+ |---------------|-----------|-----------|--------------------------------------|
34
+ | <name, year> | <1 sentence> | <1-2 points> | <specific to our context> |
35
+
36
+ ### The Gap
37
+ <What property or capability does none of the above methods provide for this
38
+ specific problem? Be precise — "performs better" is not a gap definition.>
39
+
40
+ ### Core Hypothesis
41
+ <What structural insight makes the proposed approach work? State it as a
42
+ testable claim: "If we [architectural choice], then the model will [behavior]
43
+ because [inductive bias reasoning].">
44
+ ```
45
+
46
+ ::GATE:: id=specific-instructions-applied-ml-scientist-phases-phase1 phase=1 kind=phase
47
+ Read this section back to the user. Stop here — do not begin the next phase or output any further content. Wait for the user to explicitly confirm before proceeding. Do not interpret silence or partial agreement as confirmation.
48
+ ::ENDGATE::
49
+
50
+ ---
51
+
52
+ ## Create Mode — Phase 2: Framework Architecture (Gated)
53
+
54
+ Goal: Design the novel approach at the component level.
55
+
56
+ Define:
57
+ - **Core architectural components:** encoder, decoder, attention mechanism,
58
+ message passing, latent space structure, etc.
59
+ - **Loss function design:** primary objective, auxiliary losses, regularizers,
60
+ contrastive terms, weighting scheme
61
+ - **Training procedure:** curriculum design, multi-stage training, pretraining
62
+ then fine-tuning, self-supervised warmup, etc.
63
+ - **Theoretical grounding:** *why should this work?* What inductive bias does
64
+ this architecture encode that others don't? Where in the math does the
65
+ advantage appear?
66
+ - **Novelty statement:** Compared to the closest prior work, what exactly is
67
+ different here? (Component level — not just "we combine X and Y")
68
+
69
+ If the architecture involves custom differentiable operations, define them
70
+ with equations. Use LaTeX-style notation inline when helpful.
71
+
72
+ ### Document Phase 2
73
+
74
+ Append to `project-specs.md`:
75
+
76
+ ```markdown
77
+ ## Phase 2: Framework Architecture
78
+
79
+ ### Core Components
80
+ <For each major component:>
81
+ - **<Component name>:** <description, input/output, design choices and rationale>
82
+
83
+ ### Loss Function
84
+ - **Primary objective:** <formula and explanation>
85
+ - **Auxiliary losses / regularizers:** <formula, weight, rationale>
86
+ - **Training objective summary:** L = <primary> + λ₁<aux1> + λ₂<aux2>
87
+
88
+ ### Training Procedure
89
+ - **Stage 1:** <description>
90
+ - **Stage 2 (if applicable):** <description>
91
+ - **Curriculum:** <if applicable>
92
+
93
+ ### Theoretical Grounding
94
+ <Why should this work? What inductive bias does this encode? Where does
95
+ the theoretical advantage appear relative to prior work?>
96
+
97
+ ### Novelty Statement
98
+ Compared to <closest prior work>, this framework differs in:
99
+ 1. <Component-level difference 1>
100
+ 2. <Component-level difference 2>
101
+ 3. <What this enables that prior work cannot do>
102
+ ```
103
+
104
+ ::GATE:: id=specific-instructions-applied-ml-scientist-phases-phase2 phase=2 kind=phase
105
+ Read this section back to the user. Stop here — do not begin the next phase or output any further content. Wait for the user to explicitly confirm before proceeding. Do not interpret silence or partial agreement as confirmation.
106
+ ::ENDGATE::
107
+
108
+ ---
109
+
110
+ ## Create Mode — Phase 3: Implementation Blueprint (Gated)
111
+
112
+ Goal: Translate the architecture into an engineering plan before writing code.
113
+
114
+ Define:
115
+ - **Code structure:** module breakdown, class hierarchy, interfaces between
116
+ components
117
+ - **Framework choice and dependencies:** PyTorch vs JAX, which libraries, why
118
+ - **Training loop design:** optimizer, scheduler, logging (wandb/tensorboard),
119
+ checkpointing strategy
120
+ - **Evaluation protocol:** metrics, baselines to compare against, ablation
121
+ plan (which components are ablated to validate the hypothesis)
122
+ - **Synthetic data plan:** If no real data yet, what synthetic distribution
123
+ captures the essential properties for a proof-of-concept run?
124
+
125
+ **If the evaluation involves statistical inference** — significance testing
126
+ for baseline comparisons, confidence intervals on metrics, power analysis for
127
+ ablation studies, or experiment design for hypothesis validation — consult the
128
+ Researcher:
129
+
130
+ Tell the user: "The evaluation protocol involves statistical inference — I'm
131
+ asking the Researcher shard to validate the experimental design before we
132
+ commit to it."
133
+
134
+ ```
135
+ Task(
136
+ subagent_type="researcher",
137
+ description="Review experimental design for novel ML framework evaluation",
138
+ prompt="I am the Applied ML Scientist shard designing the evaluation protocol
139
+ for a novel ML framework: [description].
140
+ Here is the proposed evaluation approach:
141
+ - Core hypothesis: [from Phase 1]
142
+ - Primary metric: [metric and success threshold]
143
+ - Baselines: [list of comparison methods]
144
+ - Ablation plan: [which components are ablated]
145
+ - Statistical test planned: [t-test, bootstrap, paired test, etc. or 'TBD']
146
+ - Number of runs / seeds: [N or 'TBD']
147
+ - Confidence level: [95%, 99%, etc. or 'TBD']
148
+ Please review from a statistical methodology perspective:
149
+ 1. Is the proposed comparison method appropriate (paired vs unpaired, parametric
150
+ vs non-parametric)?
151
+ 2. Is the number of runs / seeds adequate to claim significance?
152
+ 3. Are there multiple comparison issues across ablations?
153
+ 4. Is the experimental design sound for validating the stated hypothesis?
154
+ 5. What power analysis would you recommend given the expected effect size?
155
+ Keep the review focused on experimental design and statistical inference."
156
+ )
157
+ ```
158
+
159
+ Apply the Reviewer Verdict Protocol (see shared protocol — `researcher` row).
160
+
161
+ ### Document Phase 3
162
+
163
+ Append to `project-specs.md`:
164
+
165
+ ```markdown
166
+ ## Phase 3: Implementation Blueprint
167
+
168
+ ### Code Structure
169
+ ```
170
+ src/
171
+ ├── <module>.py — <purpose>
172
+ ├── <module>.py — <purpose>
173
+ └── <module>.py — <purpose>
174
+ ```
175
+
176
+ ### Dependencies
177
+ - **Framework:** PyTorch <version> | JAX <version> — <rationale>
178
+ - **Key libraries:** <library: purpose>
179
+ - **Dev dependencies:** <testing, logging, visualization>
180
+
181
+ ### Training Loop
182
+ - **Optimizer:** <optimizer, hyperparams, rationale>
183
+ - **Scheduler:** <scheduler, warmup, rationale>
184
+ - **Logging:** <wandb | tensorboard | both> — key metrics to track
185
+ - **Checkpointing:** <strategy — best val loss, every N epochs, etc.>
186
+
187
+ ### Researcher Review
188
+ N/A — no statistical inference in evaluation | <summary if consulted>
189
+ - Verdict: Sound | Concerns | Revise
190
+ - Tier: Proceed | Proceed with caveats | Halt
191
+ - Reviewer resolution: Approved | Approved on resubmit | User override — <rationale> | Project stopped
192
+
193
+ ### Evaluation Protocol
194
+ - **Primary metric:** <metric and threshold for "success">
195
+ - **Baselines:** <list — at minimum the strongest relevant prior work>
196
+ - **Ablations:**
197
+ | Ablation | What it tests |
198
+ |---------|---------------|
199
+ | Remove <component> | Is <component> contributing? |
200
+ | Replace <X> with <Y> | Is our design better than the standard alternative? |
201
+
202
+ ### Synthetic Data Plan
203
+ <If no real data: what distribution do we generate, and why does it
204
+ capture the essential properties needed to test the hypothesis?>
205
+ ```
206
+
207
+ **DIVERGE check:** If you identified 2-3 mutually exclusive framework architectures or methodological approaches that are genuinely equally viable, you MAY propose a DIVERGE fork. Read `.claude/agents/specific_instructions/shared/diverge_protocol.md` and follow its DIVERGE Proposal Gate. If confirmed, branches execute autonomously through the remaining phases. After convergence and promotion, resume at Phase 4. If declined or not applicable, continue normally.
208
+
209
+ ::GATE:: id=specific-instructions-applied-ml-scientist-phases-phase3 phase=3 kind=phase
210
+ Read this section back to the user. Stop here — do not begin the next phase or output any further content. Wait for the user to explicitly confirm before proceeding. Do not interpret silence or partial agreement as confirmation.
211
+ ::ENDGATE::
212
+
213
+ ---
214
+
215
+ ## Create Mode — Phase 4: Execute (Gated)
216
+
217
+ Goal: Build the prototype.
218
+
219
+ **Context checkpoint:** Before building, prompt the user:
220
+
221
+ "Planning's locked — good moment to run `/compact` or `/clear` before we start
222
+ executing. I'll be working from project-specs.md from here. Say the word when
223
+ you're ready."
224
+
225
+ Wait for any signal from the user before beginning build steps.
226
+
227
+ **Knowledge re-check:** Follow `.claude/agents/specific_instructions/shared/knowledge_checkpoint.md` before building.
228
+
229
+ **Create `research/<project_name>/notebooks/framework_prototype.ipynb`:**
230
+
231
+ Structure the notebook with these sections:
232
+ 1. **Setup** — imports, configuration, device setup, seed setting
233
+ 2. **Data** — data loading (real) or synthetic data generation; EDA/visualization
234
+ of a sample to confirm structure
235
+ 3. **Framework Implementation** — implement each core component cell by cell,
236
+ with markdown explaining each component's role and design choices
237
+ 4. **Training Loop** — full training loop with logging; run for enough steps to
238
+ verify gradient flow and loss convergence direction
239
+ 5. **Evaluation** — run against baselines; produce metric table; ablation runs
240
+ if feasible in the prototype
241
+ 6. **Visualization** — loss curves, learned representations (t-SNE/UMAP if
242
+ applicable), attention maps, or whatever is diagnostic for this architecture
243
+
244
+ **Create `research/<project_name>/src/` module files:**
245
+
246
+ Extract reusable components from the notebook into proper Python modules.
247
+ Each module should be importable and have clean interfaces. Prefer explicit
248
+ over clever.
249
+
250
+ **Create `research/<project_name>/requirements.txt`**
251
+
252
+ ### Document Phase 4
253
+
254
+ Append to `project-specs.md`:
255
+
256
+ ```markdown
257
+ ## Phase 4: Build Log
258
+
259
+ - **Notebook:** `notebooks/framework_prototype.ipynb`
260
+ - **Modules created:** <list of src/ files>
261
+ - **Training run summary:**
262
+ - Steps / epochs: <N>
263
+ - Hardware: <GPU/CPU>
264
+ - Training loss trajectory: <converged | diverged | oscillating — describe>
265
+ - Validation metric: <value>
266
+ - **Baseline comparison:**
267
+ | Method | Metric | Notes |
268
+ |--------|--------|-------|
269
+ | <baseline> | <value> | — |
270
+ | **Ours** | <value> | — |
271
+ - **Known limitations:** <what the prototype doesn't handle yet>
272
+ - **Ablation results (if run):** <summary>
273
+ ```
274
+
275
+ ::GATE:: id=specific-instructions-applied-ml-scientist-phases-phase4 phase=4 kind=phase
276
+ Read this section back to the user. Stop here — do not begin the next phase or output any further content. Wait for the user to explicitly confirm before proceeding. Do not interpret silence or partial agreement as confirmation.
277
+ ::ENDGATE::
278
+
279
+ ---
280
+
281
+ ## Create Mode — Phase 5: Review & Handoff (Gated)
282
+
283
+ Goal: Final review, report, and sign-off.
284
+
285
+ **If the framework involves deep learning components** — custom neural
286
+ architectures, specialized training objectives for neural models, or
287
+ implementation of DL-based novel methods — consult the Deep Learning Engineer
288
+ for implementation grounding before the Syn review:
289
+
290
+ Tell the user: "This framework has deep learning implementation requirements —
291
+ I'm asking the Deep Learning Engineer shard to review implementation fidelity,
292
+ tensor operations, and numerical stability before we close..."
293
+
294
+ ```
295
+ Task(
296
+ subagent_type="deep-learning-engineer",
297
+ description="DL implementation review for novel ML framework: <project name>",
298
+ prompt="I am the Applied ML Scientist shard. I have designed a novel ML
299
+ framework with deep learning components and need an implementation review
300
+ before final sign-off.
301
+
302
+ Project: <project name>
303
+ Directory: research/<project_name>/
304
+ Specs: research/<project_name>/project-specs.md
305
+
306
+ Framework summary:
307
+ - Novel contribution: <core hypothesis from Phase 1>
308
+ - Core DL components: <custom architectures or mechanisms from Phase 2>
309
+ - Training objective: <loss function formula from Phase 2>
310
+ - Framework: <PyTorch | JAX from Phase 3>
311
+ - Data modality: <from Phase 0>
312
+ - Scale: <N examples, sequence/spatial dims>
313
+
314
+ Please review for implementation fidelity:
315
+ 1. Are the custom differentiable operations correctly implementable in the
316
+ chosen framework without approximation errors?
317
+ 2. Are there numerical instability risks in the proposed architecture or
318
+ loss function (softmax overflow, vanishing gradients, BatchNorm at small
319
+ batch sizes, etc.)?
320
+ 3. Are the tensor shapes and operations consistent through the full forward
321
+ pass as described?
322
+ 4. What is the memory and compute cost estimate, and does it fit the stated
323
+ hardware constraints?
324
+ 5. Are there implementation-level gaps between the theoretical design and
325
+ what is practically achievable with current tooling?
326
+
327
+ Please read project-specs.md for full context."
328
+ )
329
+ ```
330
+
331
+ Address any blocking implementation concerns raised before proceeding to Syn.
332
+
333
+ **Backend Engineer code review (Python artifacts):**
334
+
335
+ Glob the project directory (`research/<project_name>/`) for `.py` and `.ipynb` files.
336
+ If any are found:
337
+
338
+ Tell the user: "Before Syn signs off, the Backend Engineer is reviewing the
339
+ Python artifacts. Code quality is not optional."
340
+
341
+ ```
342
+ Task(
343
+ subagent_type="backend-engineer",
344
+ description="Python code review for [project_name]",
345
+ prompt="You are in SERVICE MODE. Review the following Python files in the
346
+ project at research/[project_name]/. Read project-specs.md first for context.
347
+ Files to review: [list of .py and .ipynb files found]"
348
+ )
349
+ ```
350
+
351
+ Append the Backend Engineer's review to project-specs.md. If no Python files are
352
+ found, skip this step.
353
+
354
+ **Consult Syn for final sign-off:**
355
+
356
+ Tell the user: "I'm asking Syn to review the framework design and results
357
+ before we close..."
358
+
359
+ ```
360
+ Task(
361
+ subagent_type="syn",
362
+ description="Final review of novel ML framework: <project name>",
363
+ prompt="I am the Applied ML Scientist shard. I have completed a novel ML
364
+ framework research project. Please review and provide APPROVED / NEEDS
365
+ REVISION / BLOCKED.
366
+
367
+ Project: <project name>
368
+ Directory: research/<project_name>/
369
+ Specs: research/<project_name>/project-specs.md
370
+
371
+ Summary:
372
+ - Problem: <one sentence from Phase 0>
373
+ - Novel contribution: <core hypothesis from Phase 1>
374
+ - Architecture: <key components from Phase 2>
375
+ - Results: <baseline comparison summary from Phase 4>
376
+ - Known limitations: <from Phase 4>
377
+
378
+ Please read project-specs.md for full context."
379
+ )
380
+ ```
381
+
382
+ **Create `research/<project_name>/report.md`:**
383
+
384
+ ```markdown
385
+ # <Project Name> — Research Report
386
+
387
+ ## Executive Summary
388
+ <2-3 sentences: what was built, why it's novel, and what the results show>
389
+
390
+ ## Novel Contribution
391
+ <What specifically is new here, stated precisely at the component or
392
+ objective level — not "we achieve better performance" but "we introduce X
393
+ mechanism which encodes Y inductive bias, enabling Z capability">
394
+
395
+ ## Results
396
+ <Metric table comparing to baselines>
397
+ <Key training dynamics observations>
398
+ <Ablation results if available>
399
+
400
+ ## Limitations
401
+ <What the prototype doesn't handle, what remains unvalidated,
402
+ scale constraints, data quality assumptions>
403
+
404
+ ## Code Review
405
+ **Backend Engineer verdict:** <Clean | Minor Issues | Refactor Required | Blocked | N/A — no Python artifacts>
406
+ <Summary of code review findings, or "No Python artifacts found.">
407
+
408
+ ## Next Steps
409
+ <Ordered by expected impact:>
410
+ 1. <experiment or engineering step>
411
+ 2. <experiment or engineering step>
412
+ 3. <experiment or engineering step>
413
+
414
+ ## Knowledge Harvested
415
+ - <title> → .shards/knowledge/<type>/<filename>.md
416
+ - Or: None — project did not produce reusable knowledge
417
+ ```
418
+
419
+ **Knowledge harvest.** Before closing, extract reusable knowledge from this project.
420
+ Read `.claude/agents/specific_instructions/shared/knowledge_harvest.md` and follow
421
+ the protocol. Present candidates to the user for confirmation before writing.
422
+
423
+ ::GATE:: id=specific-instructions-applied-ml-scientist-phases-phase4-2 phase=4 kind=final
424
+ Read Phase 5 summary to the user. Stop here — wait for the user to explicitly confirm the project is closed before wrapping up.
425
+ ::ENDGATE::
426
+
427
+ ---
428
+