@proflandrigan/shards 1.1.0
This diff represents the content of publicly available package versions that have been released to one of the supported registries. The information contained in this diff is provided for informational purposes only and reflects changes between package versions as they appear in their respective public registries.
- package/README.md +475 -0
- package/package.json +37 -0
- package/src/agents/academic.md +276 -0
- package/src/agents/ai-engineer.md +377 -0
- package/src/agents/analytics-engineer.md +364 -0
- package/src/agents/applied-ml-scientist.md +410 -0
- package/src/agents/backend-engineer.md +255 -0
- package/src/agents/bi-engineer.md +333 -0
- package/src/agents/data-analyst.md +343 -0
- package/src/agents/data-engineer.md +260 -0
- package/src/agents/data-modeller.md +386 -0
- package/src/agents/data-scientist.md +366 -0
- package/src/agents/deep-learning-engineer.md +389 -0
- package/src/agents/ml-engineer.md +424 -0
- package/src/agents/mlops-engineer.md +339 -0
- package/src/agents/researcher.md +187 -0
- package/src/agents/specific_instructions/academic/critical_review.md +263 -0
- package/src/agents/specific_instructions/academic/report.md +113 -0
- package/src/agents/specific_instructions/ai_engineer/advise.md +162 -0
- package/src/agents/specific_instructions/ai_engineer/bi_engineer_handoff.md +86 -0
- package/src/agents/specific_instructions/ai_engineer/experiment.md +471 -0
- package/src/agents/specific_instructions/ai_engineer/experiment_ui_mode.md +44 -0
- package/src/agents/specific_instructions/ai_engineer/phases/index.md +45 -0
- package/src/agents/specific_instructions/ai_engineer/phases/phase-1.md +55 -0
- package/src/agents/specific_instructions/ai_engineer/phases/phase-2.md +86 -0
- package/src/agents/specific_instructions/ai_engineer/phases/phase-3.md +96 -0
- package/src/agents/specific_instructions/ai_engineer/phases/phase-4.md +138 -0
- package/src/agents/specific_instructions/ai_engineer/phases/phase-5.md +157 -0
- package/src/agents/specific_instructions/ai_engineer/phases/phase-6.md +196 -0
- package/src/agents/specific_instructions/ai_engineer/phases/phase-7.md +313 -0
- package/src/agents/specific_instructions/ai_engineer/phases.md +1011 -0
- package/src/agents/specific_instructions/ai_engineer/prompt_lab.md +161 -0
- package/src/agents/specific_instructions/ai_engineer/prompt_lab_ui_mode.md +28 -0
- package/src/agents/specific_instructions/ai_engineer/research.md +393 -0
- package/src/agents/specific_instructions/ai_engineer/research_ui_mode.md +66 -0
- package/src/agents/specific_instructions/ai_engineer/review.md +159 -0
- package/src/agents/specific_instructions/ai_engineer/validation_checklist.md +182 -0
- package/src/agents/specific_instructions/analytics_engineer/advise.md +155 -0
- package/src/agents/specific_instructions/analytics_engineer/bi_engineer_handoff.md +91 -0
- package/src/agents/specific_instructions/analytics_engineer/data_analyst_handoff.md +84 -0
- package/src/agents/specific_instructions/analytics_engineer/deep_phases.md +818 -0
- package/src/agents/specific_instructions/analytics_engineer/phases_deep/index.md +24 -0
- package/src/agents/specific_instructions/analytics_engineer/phases_deep/phase-1.md +77 -0
- package/src/agents/specific_instructions/analytics_engineer/phases_deep/phase-2.md +106 -0
- package/src/agents/specific_instructions/analytics_engineer/phases_deep/phase-3.md +93 -0
- package/src/agents/specific_instructions/analytics_engineer/phases_deep/phase-4.md +79 -0
- package/src/agents/specific_instructions/analytics_engineer/phases_deep/phase-5.md +61 -0
- package/src/agents/specific_instructions/analytics_engineer/phases_deep/phase-6.md +45 -0
- package/src/agents/specific_instructions/analytics_engineer/phases_deep/phase-7.md +235 -0
- package/src/agents/specific_instructions/analytics_engineer/phases_deep/phase-8.md +221 -0
- package/src/agents/specific_instructions/analytics_engineer/phases_quick/index.md +19 -0
- package/src/agents/specific_instructions/analytics_engineer/phases_quick/phase-1.md +47 -0
- package/src/agents/specific_instructions/analytics_engineer/phases_quick/phase-2.md +78 -0
- package/src/agents/specific_instructions/analytics_engineer/quick_phases.md +112 -0
- package/src/agents/specific_instructions/analytics_engineer/review.md +167 -0
- package/src/agents/specific_instructions/analytics_engineer/service_mode.md +369 -0
- package/src/agents/specific_instructions/analytics_engineer/ui_mode.md +45 -0
- package/src/agents/specific_instructions/analytics_engineer/update.md +162 -0
- package/src/agents/specific_instructions/analytics_engineer/validation_checklist.md +121 -0
- package/src/agents/specific_instructions/applied_ml_scientist/advise.md +143 -0
- package/src/agents/specific_instructions/applied_ml_scientist/phases/index.md +21 -0
- package/src/agents/specific_instructions/applied_ml_scientist/phases/phase-1.md +51 -0
- package/src/agents/specific_instructions/applied_ml_scientist/phases/phase-2.md +66 -0
- package/src/agents/specific_instructions/applied_ml_scientist/phases/phase-3.md +113 -0
- package/src/agents/specific_instructions/applied_ml_scientist/phases/phase-4.md +104 -0
- package/src/agents/specific_instructions/applied_ml_scientist/phases/phase-5.md +156 -0
- package/src/agents/specific_instructions/applied_ml_scientist/phases.md +428 -0
- package/src/agents/specific_instructions/applied_ml_scientist/research.md +379 -0
- package/src/agents/specific_instructions/applied_ml_scientist/review.md +142 -0
- package/src/agents/specific_instructions/applied_ml_scientist/validation_checklist.md +136 -0
- package/src/agents/specific_instructions/backend_engineer/clean.md +149 -0
- package/src/agents/specific_instructions/backend_engineer/review.md +91 -0
- package/src/agents/specific_instructions/backend_engineer/review_checklist.md +54 -0
- package/src/agents/specific_instructions/backend_engineer/service_mode.md +67 -0
- package/src/agents/specific_instructions/bi_engineer/advise.md +137 -0
- package/src/agents/specific_instructions/bi_engineer/data_analyst_handoff.md +77 -0
- package/src/agents/specific_instructions/bi_engineer/incoming_handoff.md +45 -0
- package/src/agents/specific_instructions/bi_engineer/phases/index.md +20 -0
- package/src/agents/specific_instructions/bi_engineer/phases/phase-1.md +164 -0
- package/src/agents/specific_instructions/bi_engineer/phases/phase-2.md +92 -0
- package/src/agents/specific_instructions/bi_engineer/phases/phase-3.md +121 -0
- package/src/agents/specific_instructions/bi_engineer/phases/phase-4.md +106 -0
- package/src/agents/specific_instructions/bi_engineer/phases.md +451 -0
- package/src/agents/specific_instructions/bi_engineer/review.md +166 -0
- package/src/agents/specific_instructions/bi_engineer/update.md +147 -0
- package/src/agents/specific_instructions/bi_engineer/validation_checklist.md +124 -0
- package/src/agents/specific_instructions/data_analyst/advise.md +138 -0
- package/src/agents/specific_instructions/data_analyst/explain.md +221 -0
- package/src/agents/specific_instructions/data_analyst/incoming_handoff.md +40 -0
- package/src/agents/specific_instructions/data_analyst/phases/index.md +20 -0
- package/src/agents/specific_instructions/data_analyst/phases/phase-1.md +159 -0
- package/src/agents/specific_instructions/data_analyst/phases/phase-2.md +112 -0
- package/src/agents/specific_instructions/data_analyst/phases/phase-3.md +265 -0
- package/src/agents/specific_instructions/data_analyst/phases/phase-4.md +100 -0
- package/src/agents/specific_instructions/data_analyst/phases.md +501 -0
- package/src/agents/specific_instructions/data_analyst/review.md +138 -0
- package/src/agents/specific_instructions/data_analyst/ui_mode.md +26 -0
- package/src/agents/specific_instructions/data_analyst/update.md +144 -0
- package/src/agents/specific_instructions/data_analyst/validation_checklist.md +95 -0
- package/src/agents/specific_instructions/data_engineer/advise.md +137 -0
- package/src/agents/specific_instructions/data_engineer/phases.md +466 -0
- package/src/agents/specific_instructions/data_engineer/phases_deep/index.md +23 -0
- package/src/agents/specific_instructions/data_engineer/phases_deep/phase-1.md +49 -0
- package/src/agents/specific_instructions/data_engineer/phases_deep/phase-2.md +93 -0
- package/src/agents/specific_instructions/data_engineer/phases_deep/phase-3.md +55 -0
- package/src/agents/specific_instructions/data_engineer/phases_deep/phase-4.md +48 -0
- package/src/agents/specific_instructions/data_engineer/phases_deep/phase-5.md +40 -0
- package/src/agents/specific_instructions/data_engineer/phases_deep/phase-6.md +102 -0
- package/src/agents/specific_instructions/data_engineer/phases_deep/phase-7.md +87 -0
- package/src/agents/specific_instructions/data_engineer/phases_quick/index.md +19 -0
- package/src/agents/specific_instructions/data_engineer/phases_quick/phase-1.md +45 -0
- package/src/agents/specific_instructions/data_engineer/phases_quick/phase-2.md +54 -0
- package/src/agents/specific_instructions/data_engineer/review.md +135 -0
- package/src/agents/specific_instructions/data_engineer/validation_checklist.md +136 -0
- package/src/agents/specific_instructions/data_modeller/advise.md +137 -0
- package/src/agents/specific_instructions/data_modeller/phases.md +581 -0
- package/src/agents/specific_instructions/data_modeller/phases_deep/index.md +23 -0
- package/src/agents/specific_instructions/data_modeller/phases_deep/phase-1.md +52 -0
- package/src/agents/specific_instructions/data_modeller/phases_deep/phase-2.md +113 -0
- package/src/agents/specific_instructions/data_modeller/phases_deep/phase-3.md +47 -0
- package/src/agents/specific_instructions/data_modeller/phases_deep/phase-4.md +51 -0
- package/src/agents/specific_instructions/data_modeller/phases_deep/phase-5.md +45 -0
- package/src/agents/specific_instructions/data_modeller/phases_deep/phase-6.md +105 -0
- package/src/agents/specific_instructions/data_modeller/phases_deep/phase-7.md +136 -0
- package/src/agents/specific_instructions/data_modeller/phases_quick/index.md +19 -0
- package/src/agents/specific_instructions/data_modeller/phases_quick/phase-1.md +47 -0
- package/src/agents/specific_instructions/data_modeller/phases_quick/phase-2.md +65 -0
- package/src/agents/specific_instructions/data_modeller/review.md +141 -0
- package/src/agents/specific_instructions/data_modeller/service_mode.md +218 -0
- package/src/agents/specific_instructions/data_modeller/validation_checklist.md +125 -0
- package/src/agents/specific_instructions/data_scientist/advise.md +158 -0
- package/src/agents/specific_instructions/data_scientist/bi_engineer_handoff.md +63 -0
- package/src/agents/specific_instructions/data_scientist/experiment.md +482 -0
- package/src/agents/specific_instructions/data_scientist/experiment_ui_mode.md +44 -0
- package/src/agents/specific_instructions/data_scientist/explain.md +247 -0
- package/src/agents/specific_instructions/data_scientist/greenfield_data.md +35 -0
- package/src/agents/specific_instructions/data_scientist/ml_engineer_handoff.md +52 -0
- package/src/agents/specific_instructions/data_scientist/notebook_walkthrough.md +76 -0
- package/src/agents/specific_instructions/data_scientist/phases/index.md +24 -0
- package/src/agents/specific_instructions/data_scientist/phases/phase-1.md +45 -0
- package/src/agents/specific_instructions/data_scientist/phases/phase-2.md +67 -0
- package/src/agents/specific_instructions/data_scientist/phases/phase-3.md +89 -0
- package/src/agents/specific_instructions/data_scientist/phases/phase-4.md +143 -0
- package/src/agents/specific_instructions/data_scientist/phases/phase-5.md +71 -0
- package/src/agents/specific_instructions/data_scientist/phases/phase-6.md +239 -0
- package/src/agents/specific_instructions/data_scientist/phases/phase-7.md +207 -0
- package/src/agents/specific_instructions/data_scientist/phases.md +651 -0
- package/src/agents/specific_instructions/data_scientist/research.md +345 -0
- package/src/agents/specific_instructions/data_scientist/research_ui_mode.md +52 -0
- package/src/agents/specific_instructions/data_scientist/review.md +136 -0
- package/src/agents/specific_instructions/data_scientist/service_mode.md +247 -0
- package/src/agents/specific_instructions/data_scientist/validation_checklist.md +183 -0
- package/src/agents/specific_instructions/deep_learning_engineer/advise.md +145 -0
- package/src/agents/specific_instructions/deep_learning_engineer/phases/index.md +21 -0
- package/src/agents/specific_instructions/deep_learning_engineer/phases/phase-1.md +74 -0
- package/src/agents/specific_instructions/deep_learning_engineer/phases/phase-2.md +98 -0
- package/src/agents/specific_instructions/deep_learning_engineer/phases/phase-3.md +76 -0
- package/src/agents/specific_instructions/deep_learning_engineer/phases/phase-4.md +128 -0
- package/src/agents/specific_instructions/deep_learning_engineer/phases/phase-5.md +292 -0
- package/src/agents/specific_instructions/deep_learning_engineer/phases.md +567 -0
- package/src/agents/specific_instructions/deep_learning_engineer/research.md +389 -0
- package/src/agents/specific_instructions/deep_learning_engineer/review.md +155 -0
- package/src/agents/specific_instructions/deep_learning_engineer/validation_checklist.md +147 -0
- package/src/agents/specific_instructions/ml_engineer/advise.md +174 -0
- package/src/agents/specific_instructions/ml_engineer/bi_engineer_handoff.md +71 -0
- package/src/agents/specific_instructions/ml_engineer/experiment.md +474 -0
- package/src/agents/specific_instructions/ml_engineer/experiment_ui_mode.md +44 -0
- package/src/agents/specific_instructions/ml_engineer/notebook_walkthrough.md +75 -0
- package/src/agents/specific_instructions/ml_engineer/phases/index.md +25 -0
- package/src/agents/specific_instructions/ml_engineer/phases/phase-1.md +49 -0
- package/src/agents/specific_instructions/ml_engineer/phases/phase-2.md +75 -0
- package/src/agents/specific_instructions/ml_engineer/phases/phase-3.md +124 -0
- package/src/agents/specific_instructions/ml_engineer/phases/phase-4.md +279 -0
- package/src/agents/specific_instructions/ml_engineer/phases/phase-5.md +160 -0
- package/src/agents/specific_instructions/ml_engineer/phases/phase-6-5.md +170 -0
- package/src/agents/specific_instructions/ml_engineer/phases/phase-6.md +295 -0
- package/src/agents/specific_instructions/ml_engineer/phases/phase-7.md +337 -0
- package/src/agents/specific_instructions/ml_engineer/phases.md +1068 -0
- package/src/agents/specific_instructions/ml_engineer/research.md +437 -0
- package/src/agents/specific_instructions/ml_engineer/research_ui_mode.md +71 -0
- package/src/agents/specific_instructions/ml_engineer/review.md +187 -0
- package/src/agents/specific_instructions/ml_engineer/service_mode.md +273 -0
- package/src/agents/specific_instructions/ml_engineer/validation_checklist.md +185 -0
- package/src/agents/specific_instructions/mlops_engineer/advise.md +139 -0
- package/src/agents/specific_instructions/mlops_engineer/phases/index.md +23 -0
- package/src/agents/specific_instructions/mlops_engineer/phases/phase-1.md +52 -0
- package/src/agents/specific_instructions/mlops_engineer/phases/phase-2.md +86 -0
- package/src/agents/specific_instructions/mlops_engineer/phases/phase-3.md +105 -0
- package/src/agents/specific_instructions/mlops_engineer/phases/phase-4.md +128 -0
- package/src/agents/specific_instructions/mlops_engineer/phases/phase-5.md +106 -0
- package/src/agents/specific_instructions/mlops_engineer/phases/phase-6.md +128 -0
- package/src/agents/specific_instructions/mlops_engineer/phases/phase-7.md +144 -0
- package/src/agents/specific_instructions/mlops_engineer/phases.md +671 -0
- package/src/agents/specific_instructions/mlops_engineer/review.md +164 -0
- package/src/agents/specific_instructions/mlops_engineer/service_mode.md +81 -0
- package/src/agents/specific_instructions/mlops_engineer/validation_checklist.md +151 -0
- package/src/agents/specific_instructions/researcher/critical_review.md +292 -0
- package/src/agents/specific_instructions/researcher/review_checklist.md +67 -0
- package/src/agents/specific_instructions/researcher/service_mode.md +224 -0
- package/src/agents/specific_instructions/shared/auto_verify_mode.md +141 -0
- package/src/agents/specific_instructions/shared/autonomous_research.md +1289 -0
- package/src/agents/specific_instructions/shared/behavioral_rules.md +36 -0
- package/src/agents/specific_instructions/shared/diverge_protocol.md +387 -0
- package/src/agents/specific_instructions/shared/engineering_guidelines.md +136 -0
- package/src/agents/specific_instructions/shared/experiment_versioning.md +184 -0
- package/src/agents/specific_instructions/shared/goal_mode.md +187 -0
- package/src/agents/specific_instructions/shared/incremental_testing.md +139 -0
- package/src/agents/specific_instructions/shared/intent_discovery.md +223 -0
- package/src/agents/specific_instructions/shared/join_path_protocol.md +168 -0
- package/src/agents/specific_instructions/shared/knowledge_checkpoint.md +83 -0
- package/src/agents/specific_instructions/shared/knowledge_harvest.md +220 -0
- package/src/agents/specific_instructions/shared/knowledge_retrieval.md +100 -0
- package/src/agents/specific_instructions/shared/notebook_walkthrough_protocol.md +367 -0
- package/src/agents/specific_instructions/shared/reviewer_verdict_protocol.md +74 -0
- package/src/agents/specific_instructions/shared/swarm_protocol.md +97 -0
- package/src/agents/specific_instructions/shared/validation_protocol.md +139 -0
- package/src/agents/specific_instructions/syn/arbiter.md +140 -0
- package/src/agents/specific_instructions/syn/brainstorm.md +550 -0
- package/src/agents/specific_instructions/syn/code_review.md +232 -0
- package/src/agents/specific_instructions/syn/diff.md +239 -0
- package/src/agents/specific_instructions/syn/final_review.md +65 -0
- package/src/agents/specific_instructions/syn/fixer.md +240 -0
- package/src/agents/specific_instructions/syn/free_form.md +130 -0
- package/src/agents/specific_instructions/syn/knowledge.md +468 -0
- package/src/agents/specific_instructions/syn/notebook_walkthrough.md +78 -0
- package/src/agents/specific_instructions/syn/panel_review.md +634 -0
- package/src/agents/specific_instructions/syn/pm.md +453 -0
- package/src/agents/specific_instructions/syn/pr_review.md +255 -0
- package/src/agents/specific_instructions/syn/slides.md +417 -0
- package/src/agents/syn.md +729 -0
- package/src/commands/academic.md +41 -0
- package/src/commands/ai-engineer.md +45 -0
- package/src/commands/analytics-engineer.md +48 -0
- package/src/commands/applied-ml-scientist.md +45 -0
- package/src/commands/backend-engineer.md +35 -0
- package/src/commands/bi-engineer.md +40 -0
- package/src/commands/brainstorm.md +24 -0
- package/src/commands/data-analyst.md +38 -0
- package/src/commands/data-engineer.md +37 -0
- package/src/commands/data-modeller.md +38 -0
- package/src/commands/data-scientist.md +38 -0
- package/src/commands/deep-learning-engineer.md +47 -0
- package/src/commands/end.md +49 -0
- package/src/commands/knowledge.md +24 -0
- package/src/commands/ml-engineer.md +42 -0
- package/src/commands/mlops-engineer.md +47 -0
- package/src/commands/notebook-walkthrough.md +58 -0
- package/src/commands/researcher.md +40 -0
- package/src/commands/resume.md +57 -0
- package/src/commands/review-pr.md +26 -0
- package/src/commands/shards-guide.md +41 -0
- package/src/commands/shards-ui.md +32 -0
- package/src/commands/shards.md +41 -0
- package/src/docs/01-getting-started/concepts.md +109 -0
- package/src/docs/01-getting-started/first-session.md +79 -0
- package/src/docs/01-getting-started/install.md +61 -0
- package/src/docs/02-agents/academic.md +71 -0
- package/src/docs/02-agents/ai-engineer.md +78 -0
- package/src/docs/02-agents/analytics-engineer.md +58 -0
- package/src/docs/02-agents/applied-ml-scientist.md +59 -0
- package/src/docs/02-agents/backend-engineer.md +58 -0
- package/src/docs/02-agents/bi-engineer.md +65 -0
- package/src/docs/02-agents/data-analyst.md +67 -0
- package/src/docs/02-agents/data-engineer.md +57 -0
- package/src/docs/02-agents/data-modeller.md +51 -0
- package/src/docs/02-agents/data-scientist.md +78 -0
- package/src/docs/02-agents/deep-learning-engineer.md +64 -0
- package/src/docs/02-agents/ml-engineer.md +80 -0
- package/src/docs/02-agents/mlops-engineer.md +59 -0
- package/src/docs/02-agents/overview.md +62 -0
- package/src/docs/02-agents/researcher.md +73 -0
- package/src/docs/02-agents/syn.md +88 -0
- package/src/docs/03-protocols/auto-verify.md +82 -0
- package/src/docs/03-protocols/autonomous-research.md +59 -0
- package/src/docs/03-protocols/behavioral-rules.md +35 -0
- package/src/docs/03-protocols/diverge.md +50 -0
- package/src/docs/03-protocols/engineering-guidelines.md +56 -0
- package/src/docs/03-protocols/experiment-versioning.md +38 -0
- package/src/docs/03-protocols/gate-pattern.md +65 -0
- package/src/docs/03-protocols/incremental-testing.md +68 -0
- package/src/docs/03-protocols/join-path.md +46 -0
- package/src/docs/03-protocols/knowledge-ledger.md +70 -0
- package/src/docs/03-protocols/reviewer-verdicts.md +39 -0
- package/src/docs/03-protocols/swarm.md +40 -0
- package/src/docs/03-protocols/validation.md +174 -0
- package/src/docs/04-ui/activity-bar.md +70 -0
- package/src/docs/04-ui/chat-pane.md +80 -0
- package/src/docs/04-ui/code-intel.md +62 -0
- package/src/docs/04-ui/file-editing.md +61 -0
- package/src/docs/04-ui/git.md +54 -0
- package/src/docs/04-ui/keybindings.md +79 -0
- package/src/docs/04-ui/knowledge-map.md +76 -0
- package/src/docs/04-ui/overview.md +93 -0
- package/src/docs/04-ui/panels.md +49 -0
- package/src/docs/04-ui/pinboard-selection.md +66 -0
- package/src/docs/04-ui/quick-open-palette.md +56 -0
- package/src/docs/04-ui/sessions.md +81 -0
- package/src/docs/04-ui/settings-permissions.md +56 -0
- package/src/docs/05-commands/reference.md +59 -0
- package/src/docs/06-outputs/directory-map.md +116 -0
- package/src/docs/07-workflows/ai-eval-first.md +57 -0
- package/src/docs/07-workflows/deep-study-to-production.md +76 -0
- package/src/docs/07-workflows/diverge-exploration.md +77 -0
- package/src/docs/07-workflows/quick-analysis.md +45 -0
- package/src/docs/08-integrations/claude-code-auto-mode.md +191 -0
- package/src/docs/08-integrations/google-slides.md +175 -0
- package/src/docs/README.md +30 -0
- package/src/docs/manifest.json +108 -0
- package/src/templates/analysis-template.md +20 -0
- package/src/templates/branch-report.md +46 -0
- package/src/templates/diff-report.md +88 -0
- package/src/templates/knowledge-index.md +7 -0
- package/src/templates/model-card-schema.json +186 -0
- package/src/templates/model-card-schema.md +88 -0
- package/src/templates/model-card.md +124 -0
- package/src/templates/project-plan.md +47 -0
- package/src/templates/project-specs.md +81 -0
- package/src/templates/report-template.md +43 -0
- package/src/templates/study-template.md +25 -0
- package/src/ui/cc-readonly.js +181 -0
- package/src/ui/chat-session.js +466 -0
- package/src/ui/css/base.css +136 -0
- package/src/ui/css/brainstorm.css +525 -0
- package/src/ui/css/chat.css +1405 -0
- package/src/ui/css/editor.css +546 -0
- package/src/ui/css/eval-dashboard.css +157 -0
- package/src/ui/css/experiment.css +237 -0
- package/src/ui/css/guide.css +186 -0
- package/src/ui/css/knowledge-map.css +383 -0
- package/src/ui/css/layout.css +431 -0
- package/src/ui/css/model-card.css +161 -0
- package/src/ui/css/notebook-walkthrough.css +271 -0
- package/src/ui/css/pr-review.css +403 -0
- package/src/ui/css/prompt-lab.css +325 -0
- package/src/ui/css/sessions.css +258 -0
- package/src/ui/css/sidebar.css +661 -0
- package/src/ui/css/terminal.css +113 -0
- package/src/ui/css/theme-light.css +542 -0
- package/src/ui/index.html +389 -0
- package/src/ui/js/agents.js +32 -0
- package/src/ui/js/bookmarks.js +230 -0
- package/src/ui/js/chat.js +1776 -0
- package/src/ui/js/code-intel.js +328 -0
- package/src/ui/js/command-palette.js +142 -0
- package/src/ui/js/events.js +591 -0
- package/src/ui/js/explorer.js +317 -0
- package/src/ui/js/file-view.js +477 -0
- package/src/ui/js/git.js +536 -0
- package/src/ui/js/guide.js +198 -0
- package/src/ui/js/hud.js +75 -0
- package/src/ui/js/init.js +351 -0
- package/src/ui/js/knowledge-map.js +906 -0
- package/src/ui/js/markdown.js +114 -0
- package/src/ui/js/monaco.js +164 -0
- package/src/ui/js/notebook-walkthrough.js +272 -0
- package/src/ui/js/notebook.js +448 -0
- package/src/ui/js/panels.js +2681 -0
- package/src/ui/js/pinboard.js +186 -0
- package/src/ui/js/quick-open.js +164 -0
- package/src/ui/js/selection-context.js +131 -0
- package/src/ui/js/sessions.js +256 -0
- package/src/ui/js/settings.js +476 -0
- package/src/ui/js/split-view.js +82 -0
- package/src/ui/js/state.js +343 -0
- package/src/ui/js/table.js +161 -0
- package/src/ui/js/tabs.js +284 -0
- package/src/ui/js/tabular.js +125 -0
- package/src/ui/js/terminal.js +354 -0
- package/src/ui/js/timeline.js +137 -0
- package/src/ui/js/utils.js +293 -0
- package/src/ui/notebook-kernel.py +790 -0
- package/src/ui/open-browser.js +55 -0
- package/src/ui/permission-pattern.js +42 -0
- package/src/ui/relay.js +513 -0
- package/src/ui/server.js +3072 -0
- package/src/ui/session-index.js +225 -0
- package/src/ui/shards_icon.png +0 -0
- package/src/ui/spawn-server.js +41 -0
- package/src/ui/symbol-index.js +813 -0
- package/src/ui/ui-push.js +177 -0
- package/tools/gate-hook/VALIDATION_SPEC.md +273 -0
- package/tools/gate-hook/__tests__/auto-verify.test.js +343 -0
- package/tools/gate-hook/auto-allowlist.js +179 -0
- package/tools/gate-hook/auto-state.js +68 -0
- package/tools/gate-hook/classify.js +21 -0
- package/tools/gate-hook/log.js +57 -0
- package/tools/gate-hook/parser.js +205 -0
- package/tools/gate-hook/sql-guard.js +230 -0
- package/tools/gate-hook/state.js +170 -0
- package/tools/gate-hook/sweep.js +139 -0
- package/tools/gate-hook/transcript.js +45 -0
- package/tools/gate-hook/validation.js +321 -0
- package/tools/gate-hook.js +475 -0
- package/tools/install.js +914 -0
- package/tools/shards-gates.js +311 -0
- package/tools/shards-sessions.js +261 -0
- package/tools/shards-ui.js +377 -0
|
@@ -0,0 +1,104 @@
|
|
|
1
|
+
> **Previous:** phase-3.md confirmed
|
|
2
|
+
> **Next:** phase-5.md (read only after this phase's gate is confirmed)
|
|
3
|
+
|
|
4
|
+
---
|
|
5
|
+
|
|
6
|
+
## Create Mode — Phase 4: Execute (Gated)
|
|
7
|
+
|
|
8
|
+
Goal: Build the prototype.
|
|
9
|
+
|
|
10
|
+
**Context checkpoint:** Before building, prompt the user:
|
|
11
|
+
|
|
12
|
+
"Planning's locked — good moment to run `/compact` or `/clear` before we start
|
|
13
|
+
executing. I'll be working from project-specs.md from here. Say the word when
|
|
14
|
+
you're ready."
|
|
15
|
+
|
|
16
|
+
Wait for any signal from the user before beginning build steps.
|
|
17
|
+
|
|
18
|
+
**Knowledge re-check:** Follow `.claude/agents/specific_instructions/shared/knowledge_checkpoint.md` before building.
|
|
19
|
+
|
|
20
|
+
### Incremental testing — checkpoint gates between components
|
|
21
|
+
|
|
22
|
+
Follow `.claude/agents/specific_instructions/shared/incremental_testing.md` during this build. Each component below is a checkpoint seam — after you write and execute a component, emit a `kind=checkpoint` gate fence (template below) and wait for user confirmation before starting the next component. Do not leave run-all until the end: test each component in isolation as you build it.
|
|
23
|
+
|
|
24
|
+
Checkpoint gate fence — emit exactly this shape. Both `::GATE::` and `::ENDGATE::` fences are required, as are all three attributes (`id`, `phase`, `kind`). No prose outside the fence.
|
|
25
|
+
|
|
26
|
+
```
|
|
27
|
+
::GATE:: id=<agent-name>-phase-<N>-checkpoint-<component> phase=<N> kind=checkpoint
|
|
28
|
+
Component: <human-readable name>
|
|
29
|
+
Test command: <exact command you ran>
|
|
30
|
+
Evidence:
|
|
31
|
+
- <measured fact 1, e.g. "df.shape = (48211, 47)">
|
|
32
|
+
- <measured fact 2, e.g. "null rate on join key = 0.00%">
|
|
33
|
+
- <measured fact 3, e.g. "sample head matches expected schema">
|
|
34
|
+
Status: PASS | FAIL — <one-line summary>
|
|
35
|
+
Next: <what you'll build after this is confirmed>
|
|
36
|
+
Stop here — await explicit confirmation before writing the next component.
|
|
37
|
+
::ENDGATE::
|
|
38
|
+
```
|
|
39
|
+
|
|
40
|
+
Expected checkpoint gate IDs for this phase (emit in order as you build):
|
|
41
|
+
|
|
42
|
+
- `applied-ml-scientist-phase-4-checkpoint-data` — data / synthetic-data cell produces expected shape; a sample inspection confirms structure.
|
|
43
|
+
- `applied-ml-scientist-phase-4-checkpoint-components` — each framework component forward-passes on dummy input with correct output shape and dtype.
|
|
44
|
+
- `applied-ml-scientist-phase-4-checkpoint-smoke-train` — training loop runs for 10-50 steps on a tiny batch; loss decreases (not flat, not NaN).
|
|
45
|
+
- `applied-ml-scientist-phase-4-checkpoint-full-train` — full training completes; loss curve plotted; convergence direction matches prediction.
|
|
46
|
+
- `applied-ml-scientist-phase-4-checkpoint-eval` — evaluation vs baselines runs; metric table coherent; ablation (if feasible) logged.
|
|
47
|
+
|
|
48
|
+
The hook blocks all non-read tools while a checkpoint is open. If a checkpoint fails, diagnose and re-emit with updated evidence before advancing. Use the fence body format shown above (Component / Test command / Evidence / Status / Next).
|
|
49
|
+
|
|
50
|
+
**Create `research/<project_name>/notebooks/framework_prototype.ipynb`:**
|
|
51
|
+
|
|
52
|
+
Structure the notebook with these sections:
|
|
53
|
+
1. **Setup** — imports, configuration, device setup, seed setting
|
|
54
|
+
2. **Data** — data loading (real) or synthetic data generation; EDA/visualization
|
|
55
|
+
of a sample to confirm structure
|
|
56
|
+
3. **Framework Implementation** — implement each core component cell by cell,
|
|
57
|
+
with markdown explaining each component's role and design choices
|
|
58
|
+
4. **Training Loop** — full training loop with logging; run for enough steps to
|
|
59
|
+
verify gradient flow and loss convergence direction
|
|
60
|
+
5. **Evaluation** — run against baselines; produce metric table; ablation runs
|
|
61
|
+
if feasible in the prototype
|
|
62
|
+
6. **Visualization** — loss curves, learned representations (t-SNE/UMAP if
|
|
63
|
+
applicable), attention maps, or whatever is diagnostic for this architecture
|
|
64
|
+
|
|
65
|
+
**Create `research/<project_name>/src/` module files:**
|
|
66
|
+
|
|
67
|
+
Extract reusable components from the notebook into proper Python modules.
|
|
68
|
+
Each module should be importable and have clean interfaces. Prefer explicit
|
|
69
|
+
over clever.
|
|
70
|
+
|
|
71
|
+
**Create `research/<project_name>/requirements.txt`**
|
|
72
|
+
|
|
73
|
+
### Document Phase 4
|
|
74
|
+
|
|
75
|
+
Append to `project-specs.md`:
|
|
76
|
+
|
|
77
|
+
```markdown
|
|
78
|
+
## Phase 4: Build Log
|
|
79
|
+
|
|
80
|
+
- **Notebook:** `notebooks/framework_prototype.ipynb`
|
|
81
|
+
- **Modules created:** <list of src/ files>
|
|
82
|
+
- **Training run summary:**
|
|
83
|
+
- Steps / epochs: <N>
|
|
84
|
+
- Hardware: <GPU/CPU>
|
|
85
|
+
- Training loss trajectory: <converged | diverged | oscillating — describe>
|
|
86
|
+
- Validation metric: <value>
|
|
87
|
+
- **Baseline comparison:**
|
|
88
|
+
| Method | Metric | Notes |
|
|
89
|
+
|--------|--------|-------|
|
|
90
|
+
| <baseline> | <value> | — |
|
|
91
|
+
| **Ours** | <value> | — |
|
|
92
|
+
- **Known limitations:** <what the prototype doesn't handle yet>
|
|
93
|
+
- **Ablation results (if run):** <summary>
|
|
94
|
+
```
|
|
95
|
+
|
|
96
|
+
::GATE:: id=applied-ml-scientist-phase-4 phase=4 kind=phase validates=applied_ml_scientist
|
|
97
|
+
Read this section back to the user. Stop here — do not begin the next phase or output any further content. Wait for the user to explicitly confirm before proceeding. Do not interpret silence or partial agreement as confirmation.
|
|
98
|
+
::ENDGATE::
|
|
99
|
+
|
|
100
|
+
---
|
|
101
|
+
|
|
102
|
+
## When this gate is confirmed
|
|
103
|
+
|
|
104
|
+
Read `.claude/agents/specific_instructions/applied_ml_scientist/phases/phase-5.md` in full and follow its instructions starting from Phase 5. Do not pre-read further phase files.
|
|
@@ -0,0 +1,156 @@
|
|
|
1
|
+
> **Previous:** phase-4.md confirmed
|
|
2
|
+
> **Next:** This is the final phase — follow the Syn sign-off instructions in this phase to close the project.
|
|
3
|
+
|
|
4
|
+
---
|
|
5
|
+
|
|
6
|
+
## Create Mode — Phase 5: Review & Handoff (Gated)
|
|
7
|
+
|
|
8
|
+
Goal: Final review, report, and sign-off.
|
|
9
|
+
|
|
10
|
+
**If the framework involves deep learning components** — custom neural
|
|
11
|
+
architectures, specialized training objectives for neural models, or
|
|
12
|
+
implementation of DL-based novel methods — consult the Deep Learning Engineer
|
|
13
|
+
for implementation grounding before the Syn review:
|
|
14
|
+
|
|
15
|
+
Tell the user: "This framework has deep learning implementation requirements —
|
|
16
|
+
I'm asking the Deep Learning Engineer shard to review implementation fidelity,
|
|
17
|
+
tensor operations, and numerical stability before we close..."
|
|
18
|
+
|
|
19
|
+
```
|
|
20
|
+
Task(
|
|
21
|
+
subagent_type="deep-learning-engineer",
|
|
22
|
+
description="DL implementation review for novel ML framework: <project name>",
|
|
23
|
+
prompt="I am the Applied ML Scientist shard. I have designed a novel ML
|
|
24
|
+
framework with deep learning components and need an implementation review
|
|
25
|
+
before final sign-off.
|
|
26
|
+
|
|
27
|
+
Project: <project name>
|
|
28
|
+
Directory: research/<project_name>/
|
|
29
|
+
Specs: research/<project_name>/project-specs.md
|
|
30
|
+
|
|
31
|
+
Framework summary:
|
|
32
|
+
- Novel contribution: <core hypothesis from Phase 1>
|
|
33
|
+
- Core DL components: <custom architectures or mechanisms from Phase 2>
|
|
34
|
+
- Training objective: <loss function formula from Phase 2>
|
|
35
|
+
- Framework: <PyTorch | JAX from Phase 3>
|
|
36
|
+
- Data modality: <from Phase 0>
|
|
37
|
+
- Scale: <N examples, sequence/spatial dims>
|
|
38
|
+
|
|
39
|
+
Please review for implementation fidelity:
|
|
40
|
+
1. Are the custom differentiable operations correctly implementable in the
|
|
41
|
+
chosen framework without approximation errors?
|
|
42
|
+
2. Are there numerical instability risks in the proposed architecture or
|
|
43
|
+
loss function (softmax overflow, vanishing gradients, BatchNorm at small
|
|
44
|
+
batch sizes, etc.)?
|
|
45
|
+
3. Are the tensor shapes and operations consistent through the full forward
|
|
46
|
+
pass as described?
|
|
47
|
+
4. What is the memory and compute cost estimate, and does it fit the stated
|
|
48
|
+
hardware constraints?
|
|
49
|
+
5. Are there implementation-level gaps between the theoretical design and
|
|
50
|
+
what is practically achievable with current tooling?
|
|
51
|
+
|
|
52
|
+
Please read project-specs.md for full context."
|
|
53
|
+
)
|
|
54
|
+
```
|
|
55
|
+
|
|
56
|
+
Address any blocking implementation concerns raised before proceeding to Syn.
|
|
57
|
+
|
|
58
|
+
**Backend Engineer code review (Python artifacts):**
|
|
59
|
+
|
|
60
|
+
Glob the project directory (`research/<project_name>/`) for `.py` and `.ipynb` files.
|
|
61
|
+
If any are found:
|
|
62
|
+
|
|
63
|
+
Tell the user: "Before Syn signs off, the Backend Engineer is reviewing the
|
|
64
|
+
Python artifacts. Code quality is not optional."
|
|
65
|
+
|
|
66
|
+
```
|
|
67
|
+
Task(
|
|
68
|
+
subagent_type="backend-engineer",
|
|
69
|
+
description="Python code review for [project_name]",
|
|
70
|
+
prompt="You are in SERVICE MODE. Review the following Python files in the
|
|
71
|
+
project at research/[project_name]/. Read project-specs.md first for context.
|
|
72
|
+
Files to review: [list of .py and .ipynb files found]"
|
|
73
|
+
)
|
|
74
|
+
```
|
|
75
|
+
|
|
76
|
+
Append the Backend Engineer's review to project-specs.md. If no Python files are
|
|
77
|
+
found, skip this step.
|
|
78
|
+
|
|
79
|
+
**Consult Syn for final sign-off:**
|
|
80
|
+
|
|
81
|
+
Tell the user: "I'm asking Syn to review the framework design and results
|
|
82
|
+
before we close..."
|
|
83
|
+
|
|
84
|
+
```
|
|
85
|
+
Task(
|
|
86
|
+
subagent_type="syn",
|
|
87
|
+
description="Final review of novel ML framework: <project name>",
|
|
88
|
+
prompt="I am the Applied ML Scientist shard. I have completed a novel ML
|
|
89
|
+
framework research project. Please review and provide APPROVED / NEEDS
|
|
90
|
+
REVISION / BLOCKED.
|
|
91
|
+
|
|
92
|
+
Project: <project name>
|
|
93
|
+
Directory: research/<project_name>/
|
|
94
|
+
Specs: research/<project_name>/project-specs.md
|
|
95
|
+
|
|
96
|
+
Summary:
|
|
97
|
+
- Problem: <one sentence from Phase 0>
|
|
98
|
+
- Novel contribution: <core hypothesis from Phase 1>
|
|
99
|
+
- Architecture: <key components from Phase 2>
|
|
100
|
+
- Results: <baseline comparison summary from Phase 4>
|
|
101
|
+
- Known limitations: <from Phase 4>
|
|
102
|
+
|
|
103
|
+
Please read project-specs.md for full context."
|
|
104
|
+
)
|
|
105
|
+
```
|
|
106
|
+
|
|
107
|
+
**Create `research/<project_name>/report.md`:**
|
|
108
|
+
|
|
109
|
+
```markdown
|
|
110
|
+
# <Project Name> — Research Report
|
|
111
|
+
|
|
112
|
+
## Executive Summary
|
|
113
|
+
<2-3 sentences: what was built, why it's novel, and what the results show>
|
|
114
|
+
|
|
115
|
+
## Novel Contribution
|
|
116
|
+
<What specifically is new here, stated precisely at the component or
|
|
117
|
+
objective level — not "we achieve better performance" but "we introduce X
|
|
118
|
+
mechanism which encodes Y inductive bias, enabling Z capability">
|
|
119
|
+
|
|
120
|
+
## Results
|
|
121
|
+
<Metric table comparing to baselines>
|
|
122
|
+
<Key training dynamics observations>
|
|
123
|
+
<Ablation results if available>
|
|
124
|
+
|
|
125
|
+
## Limitations
|
|
126
|
+
<What the prototype doesn't handle, what remains unvalidated,
|
|
127
|
+
scale constraints, data quality assumptions>
|
|
128
|
+
|
|
129
|
+
## Code Review
|
|
130
|
+
**Backend Engineer verdict:** <Clean | Minor Issues | Refactor Required | Blocked | N/A — no Python artifacts>
|
|
131
|
+
<Summary of code review findings, or "No Python artifacts found.">
|
|
132
|
+
|
|
133
|
+
## Next Steps
|
|
134
|
+
<Ordered by expected impact:>
|
|
135
|
+
1. <experiment or engineering step>
|
|
136
|
+
2. <experiment or engineering step>
|
|
137
|
+
3. <experiment or engineering step>
|
|
138
|
+
|
|
139
|
+
## Knowledge Harvested
|
|
140
|
+
- <title> → .shards/knowledge/<type>/<filename>.md
|
|
141
|
+
- Or: None — project did not produce reusable knowledge
|
|
142
|
+
```
|
|
143
|
+
|
|
144
|
+
**Knowledge harvest.** Before closing, extract reusable knowledge from this project.
|
|
145
|
+
Read `.claude/agents/specific_instructions/shared/knowledge_harvest.md` and follow
|
|
146
|
+
the protocol. Present candidates to the user for confirmation before writing.
|
|
147
|
+
|
|
148
|
+
::GATE:: id=applied-ml-scientist-phase-5 phase=5 kind=final
|
|
149
|
+
Read Phase 5 summary to the user. Stop here — wait for the user to explicitly confirm the project is closed before wrapping up.
|
|
150
|
+
::ENDGATE::
|
|
151
|
+
|
|
152
|
+
---
|
|
153
|
+
|
|
154
|
+
## When this gate is confirmed
|
|
155
|
+
|
|
156
|
+
This is the final phase. Once Syn returns APPROVED sign-off, the project is complete. Do not read further files.
|
|
@@ -0,0 +1,428 @@
|
|
|
1
|
+
# Applied ML Scientist — Phased Workflow (Create Mode)
|
|
2
|
+
|
|
3
|
+
Phases 1 through 5 for the Applied ML Scientist Create Mode.
|
|
4
|
+
Phase 0 (Problem Framing) is already complete.
|
|
5
|
+
Follow every phase, gate, and documentation rule below.
|
|
6
|
+
|
|
7
|
+
---
|
|
8
|
+
|
|
9
|
+
## Create Mode — Phase 1: Research Landscape (Gated)
|
|
10
|
+
|
|
11
|
+
Goal: Map the design space. Understand what exists before defining what's novel.
|
|
12
|
+
|
|
13
|
+
1. Identify the 3-5 most relevant methods or papers from the literature for
|
|
14
|
+
this problem
|
|
15
|
+
2. For each: what does it do well, and where specifically does it break down?
|
|
16
|
+
3. Identify the gap the novel framework will fill — what property does none of
|
|
17
|
+
the existing methods have?
|
|
18
|
+
4. Articulate the core hypothesis: *what structural insight makes the new
|
|
19
|
+
approach work where others don't?*
|
|
20
|
+
|
|
21
|
+
Present findings conversationally before documenting. Ask the user if any of
|
|
22
|
+
the surveyed methods are ones they've already evaluated and ruled out.
|
|
23
|
+
|
|
24
|
+
### Document Phase 1
|
|
25
|
+
|
|
26
|
+
Append to `project-specs.md`:
|
|
27
|
+
|
|
28
|
+
```markdown
|
|
29
|
+
## Phase 1: Research Landscape
|
|
30
|
+
|
|
31
|
+
### Relevant Prior Work
|
|
32
|
+
| Method / Paper | Core Idea | Strengths | Failure Modes Relevant to Our Problem |
|
|
33
|
+
|---------------|-----------|-----------|--------------------------------------|
|
|
34
|
+
| <name, year> | <1 sentence> | <1-2 points> | <specific to our context> |
|
|
35
|
+
|
|
36
|
+
### The Gap
|
|
37
|
+
<What property or capability does none of the above methods provide for this
|
|
38
|
+
specific problem? Be precise — "performs better" is not a gap definition.>
|
|
39
|
+
|
|
40
|
+
### Core Hypothesis
|
|
41
|
+
<What structural insight makes the proposed approach work? State it as a
|
|
42
|
+
testable claim: "If we [architectural choice], then the model will [behavior]
|
|
43
|
+
because [inductive bias reasoning].">
|
|
44
|
+
```
|
|
45
|
+
|
|
46
|
+
::GATE:: id=specific-instructions-applied-ml-scientist-phases-phase1 phase=1 kind=phase
|
|
47
|
+
Read this section back to the user. Stop here — do not begin the next phase or output any further content. Wait for the user to explicitly confirm before proceeding. Do not interpret silence or partial agreement as confirmation.
|
|
48
|
+
::ENDGATE::
|
|
49
|
+
|
|
50
|
+
---
|
|
51
|
+
|
|
52
|
+
## Create Mode — Phase 2: Framework Architecture (Gated)
|
|
53
|
+
|
|
54
|
+
Goal: Design the novel approach at the component level.
|
|
55
|
+
|
|
56
|
+
Define:
|
|
57
|
+
- **Core architectural components:** encoder, decoder, attention mechanism,
|
|
58
|
+
message passing, latent space structure, etc.
|
|
59
|
+
- **Loss function design:** primary objective, auxiliary losses, regularizers,
|
|
60
|
+
contrastive terms, weighting scheme
|
|
61
|
+
- **Training procedure:** curriculum design, multi-stage training, pretraining
|
|
62
|
+
then fine-tuning, self-supervised warmup, etc.
|
|
63
|
+
- **Theoretical grounding:** *why should this work?* What inductive bias does
|
|
64
|
+
this architecture encode that others don't? Where in the math does the
|
|
65
|
+
advantage appear?
|
|
66
|
+
- **Novelty statement:** Compared to the closest prior work, what exactly is
|
|
67
|
+
different here? (Component level — not just "we combine X and Y")
|
|
68
|
+
|
|
69
|
+
If the architecture involves custom differentiable operations, define them
|
|
70
|
+
with equations. Use LaTeX-style notation inline when helpful.
|
|
71
|
+
|
|
72
|
+
### Document Phase 2
|
|
73
|
+
|
|
74
|
+
Append to `project-specs.md`:
|
|
75
|
+
|
|
76
|
+
```markdown
|
|
77
|
+
## Phase 2: Framework Architecture
|
|
78
|
+
|
|
79
|
+
### Core Components
|
|
80
|
+
<For each major component:>
|
|
81
|
+
- **<Component name>:** <description, input/output, design choices and rationale>
|
|
82
|
+
|
|
83
|
+
### Loss Function
|
|
84
|
+
- **Primary objective:** <formula and explanation>
|
|
85
|
+
- **Auxiliary losses / regularizers:** <formula, weight, rationale>
|
|
86
|
+
- **Training objective summary:** L = <primary> + λ₁<aux1> + λ₂<aux2>
|
|
87
|
+
|
|
88
|
+
### Training Procedure
|
|
89
|
+
- **Stage 1:** <description>
|
|
90
|
+
- **Stage 2 (if applicable):** <description>
|
|
91
|
+
- **Curriculum:** <if applicable>
|
|
92
|
+
|
|
93
|
+
### Theoretical Grounding
|
|
94
|
+
<Why should this work? What inductive bias does this encode? Where does
|
|
95
|
+
the theoretical advantage appear relative to prior work?>
|
|
96
|
+
|
|
97
|
+
### Novelty Statement
|
|
98
|
+
Compared to <closest prior work>, this framework differs in:
|
|
99
|
+
1. <Component-level difference 1>
|
|
100
|
+
2. <Component-level difference 2>
|
|
101
|
+
3. <What this enables that prior work cannot do>
|
|
102
|
+
```
|
|
103
|
+
|
|
104
|
+
::GATE:: id=specific-instructions-applied-ml-scientist-phases-phase2 phase=2 kind=phase
|
|
105
|
+
Read this section back to the user. Stop here — do not begin the next phase or output any further content. Wait for the user to explicitly confirm before proceeding. Do not interpret silence or partial agreement as confirmation.
|
|
106
|
+
::ENDGATE::
|
|
107
|
+
|
|
108
|
+
---
|
|
109
|
+
|
|
110
|
+
## Create Mode — Phase 3: Implementation Blueprint (Gated)
|
|
111
|
+
|
|
112
|
+
Goal: Translate the architecture into an engineering plan before writing code.
|
|
113
|
+
|
|
114
|
+
Define:
|
|
115
|
+
- **Code structure:** module breakdown, class hierarchy, interfaces between
|
|
116
|
+
components
|
|
117
|
+
- **Framework choice and dependencies:** PyTorch vs JAX, which libraries, why
|
|
118
|
+
- **Training loop design:** optimizer, scheduler, logging (wandb/tensorboard),
|
|
119
|
+
checkpointing strategy
|
|
120
|
+
- **Evaluation protocol:** metrics, baselines to compare against, ablation
|
|
121
|
+
plan (which components are ablated to validate the hypothesis)
|
|
122
|
+
- **Synthetic data plan:** If no real data yet, what synthetic distribution
|
|
123
|
+
captures the essential properties for a proof-of-concept run?
|
|
124
|
+
|
|
125
|
+
**If the evaluation involves statistical inference** — significance testing
|
|
126
|
+
for baseline comparisons, confidence intervals on metrics, power analysis for
|
|
127
|
+
ablation studies, or experiment design for hypothesis validation — consult the
|
|
128
|
+
Researcher:
|
|
129
|
+
|
|
130
|
+
Tell the user: "The evaluation protocol involves statistical inference — I'm
|
|
131
|
+
asking the Researcher shard to validate the experimental design before we
|
|
132
|
+
commit to it."
|
|
133
|
+
|
|
134
|
+
```
|
|
135
|
+
Task(
|
|
136
|
+
subagent_type="researcher",
|
|
137
|
+
description="Review experimental design for novel ML framework evaluation",
|
|
138
|
+
prompt="I am the Applied ML Scientist shard designing the evaluation protocol
|
|
139
|
+
for a novel ML framework: [description].
|
|
140
|
+
Here is the proposed evaluation approach:
|
|
141
|
+
- Core hypothesis: [from Phase 1]
|
|
142
|
+
- Primary metric: [metric and success threshold]
|
|
143
|
+
- Baselines: [list of comparison methods]
|
|
144
|
+
- Ablation plan: [which components are ablated]
|
|
145
|
+
- Statistical test planned: [t-test, bootstrap, paired test, etc. or 'TBD']
|
|
146
|
+
- Number of runs / seeds: [N or 'TBD']
|
|
147
|
+
- Confidence level: [95%, 99%, etc. or 'TBD']
|
|
148
|
+
Please review from a statistical methodology perspective:
|
|
149
|
+
1. Is the proposed comparison method appropriate (paired vs unpaired, parametric
|
|
150
|
+
vs non-parametric)?
|
|
151
|
+
2. Is the number of runs / seeds adequate to claim significance?
|
|
152
|
+
3. Are there multiple comparison issues across ablations?
|
|
153
|
+
4. Is the experimental design sound for validating the stated hypothesis?
|
|
154
|
+
5. What power analysis would you recommend given the expected effect size?
|
|
155
|
+
Keep the review focused on experimental design and statistical inference."
|
|
156
|
+
)
|
|
157
|
+
```
|
|
158
|
+
|
|
159
|
+
Apply the Reviewer Verdict Protocol (see shared protocol — `researcher` row).
|
|
160
|
+
|
|
161
|
+
### Document Phase 3
|
|
162
|
+
|
|
163
|
+
Append to `project-specs.md`:
|
|
164
|
+
|
|
165
|
+
```markdown
|
|
166
|
+
## Phase 3: Implementation Blueprint
|
|
167
|
+
|
|
168
|
+
### Code Structure
|
|
169
|
+
```
|
|
170
|
+
src/
|
|
171
|
+
├── <module>.py — <purpose>
|
|
172
|
+
├── <module>.py — <purpose>
|
|
173
|
+
└── <module>.py — <purpose>
|
|
174
|
+
```
|
|
175
|
+
|
|
176
|
+
### Dependencies
|
|
177
|
+
- **Framework:** PyTorch <version> | JAX <version> — <rationale>
|
|
178
|
+
- **Key libraries:** <library: purpose>
|
|
179
|
+
- **Dev dependencies:** <testing, logging, visualization>
|
|
180
|
+
|
|
181
|
+
### Training Loop
|
|
182
|
+
- **Optimizer:** <optimizer, hyperparams, rationale>
|
|
183
|
+
- **Scheduler:** <scheduler, warmup, rationale>
|
|
184
|
+
- **Logging:** <wandb | tensorboard | both> — key metrics to track
|
|
185
|
+
- **Checkpointing:** <strategy — best val loss, every N epochs, etc.>
|
|
186
|
+
|
|
187
|
+
### Researcher Review
|
|
188
|
+
N/A — no statistical inference in evaluation | <summary if consulted>
|
|
189
|
+
- Verdict: Sound | Concerns | Revise
|
|
190
|
+
- Tier: Proceed | Proceed with caveats | Halt
|
|
191
|
+
- Reviewer resolution: Approved | Approved on resubmit | User override — <rationale> | Project stopped
|
|
192
|
+
|
|
193
|
+
### Evaluation Protocol
|
|
194
|
+
- **Primary metric:** <metric and threshold for "success">
|
|
195
|
+
- **Baselines:** <list — at minimum the strongest relevant prior work>
|
|
196
|
+
- **Ablations:**
|
|
197
|
+
| Ablation | What it tests |
|
|
198
|
+
|---------|---------------|
|
|
199
|
+
| Remove <component> | Is <component> contributing? |
|
|
200
|
+
| Replace <X> with <Y> | Is our design better than the standard alternative? |
|
|
201
|
+
|
|
202
|
+
### Synthetic Data Plan
|
|
203
|
+
<If no real data: what distribution do we generate, and why does it
|
|
204
|
+
capture the essential properties needed to test the hypothesis?>
|
|
205
|
+
```
|
|
206
|
+
|
|
207
|
+
**DIVERGE check:** If you identified 2-3 mutually exclusive framework architectures or methodological approaches that are genuinely equally viable, you MAY propose a DIVERGE fork. Read `.claude/agents/specific_instructions/shared/diverge_protocol.md` and follow its DIVERGE Proposal Gate. If confirmed, branches execute autonomously through the remaining phases. After convergence and promotion, resume at Phase 4. If declined or not applicable, continue normally.
|
|
208
|
+
|
|
209
|
+
::GATE:: id=specific-instructions-applied-ml-scientist-phases-phase3 phase=3 kind=phase
|
|
210
|
+
Read this section back to the user. Stop here — do not begin the next phase or output any further content. Wait for the user to explicitly confirm before proceeding. Do not interpret silence or partial agreement as confirmation.
|
|
211
|
+
::ENDGATE::
|
|
212
|
+
|
|
213
|
+
---
|
|
214
|
+
|
|
215
|
+
## Create Mode — Phase 4: Execute (Gated)
|
|
216
|
+
|
|
217
|
+
Goal: Build the prototype.
|
|
218
|
+
|
|
219
|
+
**Context checkpoint:** Before building, prompt the user:
|
|
220
|
+
|
|
221
|
+
"Planning's locked — good moment to run `/compact` or `/clear` before we start
|
|
222
|
+
executing. I'll be working from project-specs.md from here. Say the word when
|
|
223
|
+
you're ready."
|
|
224
|
+
|
|
225
|
+
Wait for any signal from the user before beginning build steps.
|
|
226
|
+
|
|
227
|
+
**Knowledge re-check:** Follow `.claude/agents/specific_instructions/shared/knowledge_checkpoint.md` before building.
|
|
228
|
+
|
|
229
|
+
**Create `research/<project_name>/notebooks/framework_prototype.ipynb`:**
|
|
230
|
+
|
|
231
|
+
Structure the notebook with these sections:
|
|
232
|
+
1. **Setup** — imports, configuration, device setup, seed setting
|
|
233
|
+
2. **Data** — data loading (real) or synthetic data generation; EDA/visualization
|
|
234
|
+
of a sample to confirm structure
|
|
235
|
+
3. **Framework Implementation** — implement each core component cell by cell,
|
|
236
|
+
with markdown explaining each component's role and design choices
|
|
237
|
+
4. **Training Loop** — full training loop with logging; run for enough steps to
|
|
238
|
+
verify gradient flow and loss convergence direction
|
|
239
|
+
5. **Evaluation** — run against baselines; produce metric table; ablation runs
|
|
240
|
+
if feasible in the prototype
|
|
241
|
+
6. **Visualization** — loss curves, learned representations (t-SNE/UMAP if
|
|
242
|
+
applicable), attention maps, or whatever is diagnostic for this architecture
|
|
243
|
+
|
|
244
|
+
**Create `research/<project_name>/src/` module files:**
|
|
245
|
+
|
|
246
|
+
Extract reusable components from the notebook into proper Python modules.
|
|
247
|
+
Each module should be importable and have clean interfaces. Prefer explicit
|
|
248
|
+
over clever.
|
|
249
|
+
|
|
250
|
+
**Create `research/<project_name>/requirements.txt`**
|
|
251
|
+
|
|
252
|
+
### Document Phase 4
|
|
253
|
+
|
|
254
|
+
Append to `project-specs.md`:
|
|
255
|
+
|
|
256
|
+
```markdown
|
|
257
|
+
## Phase 4: Build Log
|
|
258
|
+
|
|
259
|
+
- **Notebook:** `notebooks/framework_prototype.ipynb`
|
|
260
|
+
- **Modules created:** <list of src/ files>
|
|
261
|
+
- **Training run summary:**
|
|
262
|
+
- Steps / epochs: <N>
|
|
263
|
+
- Hardware: <GPU/CPU>
|
|
264
|
+
- Training loss trajectory: <converged | diverged | oscillating — describe>
|
|
265
|
+
- Validation metric: <value>
|
|
266
|
+
- **Baseline comparison:**
|
|
267
|
+
| Method | Metric | Notes |
|
|
268
|
+
|--------|--------|-------|
|
|
269
|
+
| <baseline> | <value> | — |
|
|
270
|
+
| **Ours** | <value> | — |
|
|
271
|
+
- **Known limitations:** <what the prototype doesn't handle yet>
|
|
272
|
+
- **Ablation results (if run):** <summary>
|
|
273
|
+
```
|
|
274
|
+
|
|
275
|
+
::GATE:: id=specific-instructions-applied-ml-scientist-phases-phase4 phase=4 kind=phase
|
|
276
|
+
Read this section back to the user. Stop here — do not begin the next phase or output any further content. Wait for the user to explicitly confirm before proceeding. Do not interpret silence or partial agreement as confirmation.
|
|
277
|
+
::ENDGATE::
|
|
278
|
+
|
|
279
|
+
---
|
|
280
|
+
|
|
281
|
+
## Create Mode — Phase 5: Review & Handoff (Gated)
|
|
282
|
+
|
|
283
|
+
Goal: Final review, report, and sign-off.
|
|
284
|
+
|
|
285
|
+
**If the framework involves deep learning components** — custom neural
|
|
286
|
+
architectures, specialized training objectives for neural models, or
|
|
287
|
+
implementation of DL-based novel methods — consult the Deep Learning Engineer
|
|
288
|
+
for implementation grounding before the Syn review:
|
|
289
|
+
|
|
290
|
+
Tell the user: "This framework has deep learning implementation requirements —
|
|
291
|
+
I'm asking the Deep Learning Engineer shard to review implementation fidelity,
|
|
292
|
+
tensor operations, and numerical stability before we close..."
|
|
293
|
+
|
|
294
|
+
```
|
|
295
|
+
Task(
|
|
296
|
+
subagent_type="deep-learning-engineer",
|
|
297
|
+
description="DL implementation review for novel ML framework: <project name>",
|
|
298
|
+
prompt="I am the Applied ML Scientist shard. I have designed a novel ML
|
|
299
|
+
framework with deep learning components and need an implementation review
|
|
300
|
+
before final sign-off.
|
|
301
|
+
|
|
302
|
+
Project: <project name>
|
|
303
|
+
Directory: research/<project_name>/
|
|
304
|
+
Specs: research/<project_name>/project-specs.md
|
|
305
|
+
|
|
306
|
+
Framework summary:
|
|
307
|
+
- Novel contribution: <core hypothesis from Phase 1>
|
|
308
|
+
- Core DL components: <custom architectures or mechanisms from Phase 2>
|
|
309
|
+
- Training objective: <loss function formula from Phase 2>
|
|
310
|
+
- Framework: <PyTorch | JAX from Phase 3>
|
|
311
|
+
- Data modality: <from Phase 0>
|
|
312
|
+
- Scale: <N examples, sequence/spatial dims>
|
|
313
|
+
|
|
314
|
+
Please review for implementation fidelity:
|
|
315
|
+
1. Are the custom differentiable operations correctly implementable in the
|
|
316
|
+
chosen framework without approximation errors?
|
|
317
|
+
2. Are there numerical instability risks in the proposed architecture or
|
|
318
|
+
loss function (softmax overflow, vanishing gradients, BatchNorm at small
|
|
319
|
+
batch sizes, etc.)?
|
|
320
|
+
3. Are the tensor shapes and operations consistent through the full forward
|
|
321
|
+
pass as described?
|
|
322
|
+
4. What is the memory and compute cost estimate, and does it fit the stated
|
|
323
|
+
hardware constraints?
|
|
324
|
+
5. Are there implementation-level gaps between the theoretical design and
|
|
325
|
+
what is practically achievable with current tooling?
|
|
326
|
+
|
|
327
|
+
Please read project-specs.md for full context."
|
|
328
|
+
)
|
|
329
|
+
```
|
|
330
|
+
|
|
331
|
+
Address any blocking implementation concerns raised before proceeding to Syn.
|
|
332
|
+
|
|
333
|
+
**Backend Engineer code review (Python artifacts):**
|
|
334
|
+
|
|
335
|
+
Glob the project directory (`research/<project_name>/`) for `.py` and `.ipynb` files.
|
|
336
|
+
If any are found:
|
|
337
|
+
|
|
338
|
+
Tell the user: "Before Syn signs off, the Backend Engineer is reviewing the
|
|
339
|
+
Python artifacts. Code quality is not optional."
|
|
340
|
+
|
|
341
|
+
```
|
|
342
|
+
Task(
|
|
343
|
+
subagent_type="backend-engineer",
|
|
344
|
+
description="Python code review for [project_name]",
|
|
345
|
+
prompt="You are in SERVICE MODE. Review the following Python files in the
|
|
346
|
+
project at research/[project_name]/. Read project-specs.md first for context.
|
|
347
|
+
Files to review: [list of .py and .ipynb files found]"
|
|
348
|
+
)
|
|
349
|
+
```
|
|
350
|
+
|
|
351
|
+
Append the Backend Engineer's review to project-specs.md. If no Python files are
|
|
352
|
+
found, skip this step.
|
|
353
|
+
|
|
354
|
+
**Consult Syn for final sign-off:**
|
|
355
|
+
|
|
356
|
+
Tell the user: "I'm asking Syn to review the framework design and results
|
|
357
|
+
before we close..."
|
|
358
|
+
|
|
359
|
+
```
|
|
360
|
+
Task(
|
|
361
|
+
subagent_type="syn",
|
|
362
|
+
description="Final review of novel ML framework: <project name>",
|
|
363
|
+
prompt="I am the Applied ML Scientist shard. I have completed a novel ML
|
|
364
|
+
framework research project. Please review and provide APPROVED / NEEDS
|
|
365
|
+
REVISION / BLOCKED.
|
|
366
|
+
|
|
367
|
+
Project: <project name>
|
|
368
|
+
Directory: research/<project_name>/
|
|
369
|
+
Specs: research/<project_name>/project-specs.md
|
|
370
|
+
|
|
371
|
+
Summary:
|
|
372
|
+
- Problem: <one sentence from Phase 0>
|
|
373
|
+
- Novel contribution: <core hypothesis from Phase 1>
|
|
374
|
+
- Architecture: <key components from Phase 2>
|
|
375
|
+
- Results: <baseline comparison summary from Phase 4>
|
|
376
|
+
- Known limitations: <from Phase 4>
|
|
377
|
+
|
|
378
|
+
Please read project-specs.md for full context."
|
|
379
|
+
)
|
|
380
|
+
```
|
|
381
|
+
|
|
382
|
+
**Create `research/<project_name>/report.md`:**
|
|
383
|
+
|
|
384
|
+
```markdown
|
|
385
|
+
# <Project Name> — Research Report
|
|
386
|
+
|
|
387
|
+
## Executive Summary
|
|
388
|
+
<2-3 sentences: what was built, why it's novel, and what the results show>
|
|
389
|
+
|
|
390
|
+
## Novel Contribution
|
|
391
|
+
<What specifically is new here, stated precisely at the component or
|
|
392
|
+
objective level — not "we achieve better performance" but "we introduce X
|
|
393
|
+
mechanism which encodes Y inductive bias, enabling Z capability">
|
|
394
|
+
|
|
395
|
+
## Results
|
|
396
|
+
<Metric table comparing to baselines>
|
|
397
|
+
<Key training dynamics observations>
|
|
398
|
+
<Ablation results if available>
|
|
399
|
+
|
|
400
|
+
## Limitations
|
|
401
|
+
<What the prototype doesn't handle, what remains unvalidated,
|
|
402
|
+
scale constraints, data quality assumptions>
|
|
403
|
+
|
|
404
|
+
## Code Review
|
|
405
|
+
**Backend Engineer verdict:** <Clean | Minor Issues | Refactor Required | Blocked | N/A — no Python artifacts>
|
|
406
|
+
<Summary of code review findings, or "No Python artifacts found.">
|
|
407
|
+
|
|
408
|
+
## Next Steps
|
|
409
|
+
<Ordered by expected impact:>
|
|
410
|
+
1. <experiment or engineering step>
|
|
411
|
+
2. <experiment or engineering step>
|
|
412
|
+
3. <experiment or engineering step>
|
|
413
|
+
|
|
414
|
+
## Knowledge Harvested
|
|
415
|
+
- <title> → .shards/knowledge/<type>/<filename>.md
|
|
416
|
+
- Or: None — project did not produce reusable knowledge
|
|
417
|
+
```
|
|
418
|
+
|
|
419
|
+
**Knowledge harvest.** Before closing, extract reusable knowledge from this project.
|
|
420
|
+
Read `.claude/agents/specific_instructions/shared/knowledge_harvest.md` and follow
|
|
421
|
+
the protocol. Present candidates to the user for confirmation before writing.
|
|
422
|
+
|
|
423
|
+
::GATE:: id=specific-instructions-applied-ml-scientist-phases-phase4-2 phase=4 kind=final
|
|
424
|
+
Read Phase 5 summary to the user. Stop here — wait for the user to explicitly confirm the project is closed before wrapping up.
|
|
425
|
+
::ENDGATE::
|
|
426
|
+
|
|
427
|
+
---
|
|
428
|
+
|