@proflandrigan/shards 1.1.0
This diff represents the content of publicly available package versions that have been released to one of the supported registries. The information contained in this diff is provided for informational purposes only and reflects changes between package versions as they appear in their respective public registries.
- package/README.md +475 -0
- package/package.json +37 -0
- package/src/agents/academic.md +276 -0
- package/src/agents/ai-engineer.md +377 -0
- package/src/agents/analytics-engineer.md +364 -0
- package/src/agents/applied-ml-scientist.md +410 -0
- package/src/agents/backend-engineer.md +255 -0
- package/src/agents/bi-engineer.md +333 -0
- package/src/agents/data-analyst.md +343 -0
- package/src/agents/data-engineer.md +260 -0
- package/src/agents/data-modeller.md +386 -0
- package/src/agents/data-scientist.md +366 -0
- package/src/agents/deep-learning-engineer.md +389 -0
- package/src/agents/ml-engineer.md +424 -0
- package/src/agents/mlops-engineer.md +339 -0
- package/src/agents/researcher.md +187 -0
- package/src/agents/specific_instructions/academic/critical_review.md +263 -0
- package/src/agents/specific_instructions/academic/report.md +113 -0
- package/src/agents/specific_instructions/ai_engineer/advise.md +162 -0
- package/src/agents/specific_instructions/ai_engineer/bi_engineer_handoff.md +86 -0
- package/src/agents/specific_instructions/ai_engineer/experiment.md +471 -0
- package/src/agents/specific_instructions/ai_engineer/experiment_ui_mode.md +44 -0
- package/src/agents/specific_instructions/ai_engineer/phases/index.md +45 -0
- package/src/agents/specific_instructions/ai_engineer/phases/phase-1.md +55 -0
- package/src/agents/specific_instructions/ai_engineer/phases/phase-2.md +86 -0
- package/src/agents/specific_instructions/ai_engineer/phases/phase-3.md +96 -0
- package/src/agents/specific_instructions/ai_engineer/phases/phase-4.md +138 -0
- package/src/agents/specific_instructions/ai_engineer/phases/phase-5.md +157 -0
- package/src/agents/specific_instructions/ai_engineer/phases/phase-6.md +196 -0
- package/src/agents/specific_instructions/ai_engineer/phases/phase-7.md +313 -0
- package/src/agents/specific_instructions/ai_engineer/phases.md +1011 -0
- package/src/agents/specific_instructions/ai_engineer/prompt_lab.md +161 -0
- package/src/agents/specific_instructions/ai_engineer/prompt_lab_ui_mode.md +28 -0
- package/src/agents/specific_instructions/ai_engineer/research.md +393 -0
- package/src/agents/specific_instructions/ai_engineer/research_ui_mode.md +66 -0
- package/src/agents/specific_instructions/ai_engineer/review.md +159 -0
- package/src/agents/specific_instructions/ai_engineer/validation_checklist.md +182 -0
- package/src/agents/specific_instructions/analytics_engineer/advise.md +155 -0
- package/src/agents/specific_instructions/analytics_engineer/bi_engineer_handoff.md +91 -0
- package/src/agents/specific_instructions/analytics_engineer/data_analyst_handoff.md +84 -0
- package/src/agents/specific_instructions/analytics_engineer/deep_phases.md +818 -0
- package/src/agents/specific_instructions/analytics_engineer/phases_deep/index.md +24 -0
- package/src/agents/specific_instructions/analytics_engineer/phases_deep/phase-1.md +77 -0
- package/src/agents/specific_instructions/analytics_engineer/phases_deep/phase-2.md +106 -0
- package/src/agents/specific_instructions/analytics_engineer/phases_deep/phase-3.md +93 -0
- package/src/agents/specific_instructions/analytics_engineer/phases_deep/phase-4.md +79 -0
- package/src/agents/specific_instructions/analytics_engineer/phases_deep/phase-5.md +61 -0
- package/src/agents/specific_instructions/analytics_engineer/phases_deep/phase-6.md +45 -0
- package/src/agents/specific_instructions/analytics_engineer/phases_deep/phase-7.md +235 -0
- package/src/agents/specific_instructions/analytics_engineer/phases_deep/phase-8.md +221 -0
- package/src/agents/specific_instructions/analytics_engineer/phases_quick/index.md +19 -0
- package/src/agents/specific_instructions/analytics_engineer/phases_quick/phase-1.md +47 -0
- package/src/agents/specific_instructions/analytics_engineer/phases_quick/phase-2.md +78 -0
- package/src/agents/specific_instructions/analytics_engineer/quick_phases.md +112 -0
- package/src/agents/specific_instructions/analytics_engineer/review.md +167 -0
- package/src/agents/specific_instructions/analytics_engineer/service_mode.md +369 -0
- package/src/agents/specific_instructions/analytics_engineer/ui_mode.md +45 -0
- package/src/agents/specific_instructions/analytics_engineer/update.md +162 -0
- package/src/agents/specific_instructions/analytics_engineer/validation_checklist.md +121 -0
- package/src/agents/specific_instructions/applied_ml_scientist/advise.md +143 -0
- package/src/agents/specific_instructions/applied_ml_scientist/phases/index.md +21 -0
- package/src/agents/specific_instructions/applied_ml_scientist/phases/phase-1.md +51 -0
- package/src/agents/specific_instructions/applied_ml_scientist/phases/phase-2.md +66 -0
- package/src/agents/specific_instructions/applied_ml_scientist/phases/phase-3.md +113 -0
- package/src/agents/specific_instructions/applied_ml_scientist/phases/phase-4.md +104 -0
- package/src/agents/specific_instructions/applied_ml_scientist/phases/phase-5.md +156 -0
- package/src/agents/specific_instructions/applied_ml_scientist/phases.md +428 -0
- package/src/agents/specific_instructions/applied_ml_scientist/research.md +379 -0
- package/src/agents/specific_instructions/applied_ml_scientist/review.md +142 -0
- package/src/agents/specific_instructions/applied_ml_scientist/validation_checklist.md +136 -0
- package/src/agents/specific_instructions/backend_engineer/clean.md +149 -0
- package/src/agents/specific_instructions/backend_engineer/review.md +91 -0
- package/src/agents/specific_instructions/backend_engineer/review_checklist.md +54 -0
- package/src/agents/specific_instructions/backend_engineer/service_mode.md +67 -0
- package/src/agents/specific_instructions/bi_engineer/advise.md +137 -0
- package/src/agents/specific_instructions/bi_engineer/data_analyst_handoff.md +77 -0
- package/src/agents/specific_instructions/bi_engineer/incoming_handoff.md +45 -0
- package/src/agents/specific_instructions/bi_engineer/phases/index.md +20 -0
- package/src/agents/specific_instructions/bi_engineer/phases/phase-1.md +164 -0
- package/src/agents/specific_instructions/bi_engineer/phases/phase-2.md +92 -0
- package/src/agents/specific_instructions/bi_engineer/phases/phase-3.md +121 -0
- package/src/agents/specific_instructions/bi_engineer/phases/phase-4.md +106 -0
- package/src/agents/specific_instructions/bi_engineer/phases.md +451 -0
- package/src/agents/specific_instructions/bi_engineer/review.md +166 -0
- package/src/agents/specific_instructions/bi_engineer/update.md +147 -0
- package/src/agents/specific_instructions/bi_engineer/validation_checklist.md +124 -0
- package/src/agents/specific_instructions/data_analyst/advise.md +138 -0
- package/src/agents/specific_instructions/data_analyst/explain.md +221 -0
- package/src/agents/specific_instructions/data_analyst/incoming_handoff.md +40 -0
- package/src/agents/specific_instructions/data_analyst/phases/index.md +20 -0
- package/src/agents/specific_instructions/data_analyst/phases/phase-1.md +159 -0
- package/src/agents/specific_instructions/data_analyst/phases/phase-2.md +112 -0
- package/src/agents/specific_instructions/data_analyst/phases/phase-3.md +265 -0
- package/src/agents/specific_instructions/data_analyst/phases/phase-4.md +100 -0
- package/src/agents/specific_instructions/data_analyst/phases.md +501 -0
- package/src/agents/specific_instructions/data_analyst/review.md +138 -0
- package/src/agents/specific_instructions/data_analyst/ui_mode.md +26 -0
- package/src/agents/specific_instructions/data_analyst/update.md +144 -0
- package/src/agents/specific_instructions/data_analyst/validation_checklist.md +95 -0
- package/src/agents/specific_instructions/data_engineer/advise.md +137 -0
- package/src/agents/specific_instructions/data_engineer/phases.md +466 -0
- package/src/agents/specific_instructions/data_engineer/phases_deep/index.md +23 -0
- package/src/agents/specific_instructions/data_engineer/phases_deep/phase-1.md +49 -0
- package/src/agents/specific_instructions/data_engineer/phases_deep/phase-2.md +93 -0
- package/src/agents/specific_instructions/data_engineer/phases_deep/phase-3.md +55 -0
- package/src/agents/specific_instructions/data_engineer/phases_deep/phase-4.md +48 -0
- package/src/agents/specific_instructions/data_engineer/phases_deep/phase-5.md +40 -0
- package/src/agents/specific_instructions/data_engineer/phases_deep/phase-6.md +102 -0
- package/src/agents/specific_instructions/data_engineer/phases_deep/phase-7.md +87 -0
- package/src/agents/specific_instructions/data_engineer/phases_quick/index.md +19 -0
- package/src/agents/specific_instructions/data_engineer/phases_quick/phase-1.md +45 -0
- package/src/agents/specific_instructions/data_engineer/phases_quick/phase-2.md +54 -0
- package/src/agents/specific_instructions/data_engineer/review.md +135 -0
- package/src/agents/specific_instructions/data_engineer/validation_checklist.md +136 -0
- package/src/agents/specific_instructions/data_modeller/advise.md +137 -0
- package/src/agents/specific_instructions/data_modeller/phases.md +581 -0
- package/src/agents/specific_instructions/data_modeller/phases_deep/index.md +23 -0
- package/src/agents/specific_instructions/data_modeller/phases_deep/phase-1.md +52 -0
- package/src/agents/specific_instructions/data_modeller/phases_deep/phase-2.md +113 -0
- package/src/agents/specific_instructions/data_modeller/phases_deep/phase-3.md +47 -0
- package/src/agents/specific_instructions/data_modeller/phases_deep/phase-4.md +51 -0
- package/src/agents/specific_instructions/data_modeller/phases_deep/phase-5.md +45 -0
- package/src/agents/specific_instructions/data_modeller/phases_deep/phase-6.md +105 -0
- package/src/agents/specific_instructions/data_modeller/phases_deep/phase-7.md +136 -0
- package/src/agents/specific_instructions/data_modeller/phases_quick/index.md +19 -0
- package/src/agents/specific_instructions/data_modeller/phases_quick/phase-1.md +47 -0
- package/src/agents/specific_instructions/data_modeller/phases_quick/phase-2.md +65 -0
- package/src/agents/specific_instructions/data_modeller/review.md +141 -0
- package/src/agents/specific_instructions/data_modeller/service_mode.md +218 -0
- package/src/agents/specific_instructions/data_modeller/validation_checklist.md +125 -0
- package/src/agents/specific_instructions/data_scientist/advise.md +158 -0
- package/src/agents/specific_instructions/data_scientist/bi_engineer_handoff.md +63 -0
- package/src/agents/specific_instructions/data_scientist/experiment.md +482 -0
- package/src/agents/specific_instructions/data_scientist/experiment_ui_mode.md +44 -0
- package/src/agents/specific_instructions/data_scientist/explain.md +247 -0
- package/src/agents/specific_instructions/data_scientist/greenfield_data.md +35 -0
- package/src/agents/specific_instructions/data_scientist/ml_engineer_handoff.md +52 -0
- package/src/agents/specific_instructions/data_scientist/notebook_walkthrough.md +76 -0
- package/src/agents/specific_instructions/data_scientist/phases/index.md +24 -0
- package/src/agents/specific_instructions/data_scientist/phases/phase-1.md +45 -0
- package/src/agents/specific_instructions/data_scientist/phases/phase-2.md +67 -0
- package/src/agents/specific_instructions/data_scientist/phases/phase-3.md +89 -0
- package/src/agents/specific_instructions/data_scientist/phases/phase-4.md +143 -0
- package/src/agents/specific_instructions/data_scientist/phases/phase-5.md +71 -0
- package/src/agents/specific_instructions/data_scientist/phases/phase-6.md +239 -0
- package/src/agents/specific_instructions/data_scientist/phases/phase-7.md +207 -0
- package/src/agents/specific_instructions/data_scientist/phases.md +651 -0
- package/src/agents/specific_instructions/data_scientist/research.md +345 -0
- package/src/agents/specific_instructions/data_scientist/research_ui_mode.md +52 -0
- package/src/agents/specific_instructions/data_scientist/review.md +136 -0
- package/src/agents/specific_instructions/data_scientist/service_mode.md +247 -0
- package/src/agents/specific_instructions/data_scientist/validation_checklist.md +183 -0
- package/src/agents/specific_instructions/deep_learning_engineer/advise.md +145 -0
- package/src/agents/specific_instructions/deep_learning_engineer/phases/index.md +21 -0
- package/src/agents/specific_instructions/deep_learning_engineer/phases/phase-1.md +74 -0
- package/src/agents/specific_instructions/deep_learning_engineer/phases/phase-2.md +98 -0
- package/src/agents/specific_instructions/deep_learning_engineer/phases/phase-3.md +76 -0
- package/src/agents/specific_instructions/deep_learning_engineer/phases/phase-4.md +128 -0
- package/src/agents/specific_instructions/deep_learning_engineer/phases/phase-5.md +292 -0
- package/src/agents/specific_instructions/deep_learning_engineer/phases.md +567 -0
- package/src/agents/specific_instructions/deep_learning_engineer/research.md +389 -0
- package/src/agents/specific_instructions/deep_learning_engineer/review.md +155 -0
- package/src/agents/specific_instructions/deep_learning_engineer/validation_checklist.md +147 -0
- package/src/agents/specific_instructions/ml_engineer/advise.md +174 -0
- package/src/agents/specific_instructions/ml_engineer/bi_engineer_handoff.md +71 -0
- package/src/agents/specific_instructions/ml_engineer/experiment.md +474 -0
- package/src/agents/specific_instructions/ml_engineer/experiment_ui_mode.md +44 -0
- package/src/agents/specific_instructions/ml_engineer/notebook_walkthrough.md +75 -0
- package/src/agents/specific_instructions/ml_engineer/phases/index.md +25 -0
- package/src/agents/specific_instructions/ml_engineer/phases/phase-1.md +49 -0
- package/src/agents/specific_instructions/ml_engineer/phases/phase-2.md +75 -0
- package/src/agents/specific_instructions/ml_engineer/phases/phase-3.md +124 -0
- package/src/agents/specific_instructions/ml_engineer/phases/phase-4.md +279 -0
- package/src/agents/specific_instructions/ml_engineer/phases/phase-5.md +160 -0
- package/src/agents/specific_instructions/ml_engineer/phases/phase-6-5.md +170 -0
- package/src/agents/specific_instructions/ml_engineer/phases/phase-6.md +295 -0
- package/src/agents/specific_instructions/ml_engineer/phases/phase-7.md +337 -0
- package/src/agents/specific_instructions/ml_engineer/phases.md +1068 -0
- package/src/agents/specific_instructions/ml_engineer/research.md +437 -0
- package/src/agents/specific_instructions/ml_engineer/research_ui_mode.md +71 -0
- package/src/agents/specific_instructions/ml_engineer/review.md +187 -0
- package/src/agents/specific_instructions/ml_engineer/service_mode.md +273 -0
- package/src/agents/specific_instructions/ml_engineer/validation_checklist.md +185 -0
- package/src/agents/specific_instructions/mlops_engineer/advise.md +139 -0
- package/src/agents/specific_instructions/mlops_engineer/phases/index.md +23 -0
- package/src/agents/specific_instructions/mlops_engineer/phases/phase-1.md +52 -0
- package/src/agents/specific_instructions/mlops_engineer/phases/phase-2.md +86 -0
- package/src/agents/specific_instructions/mlops_engineer/phases/phase-3.md +105 -0
- package/src/agents/specific_instructions/mlops_engineer/phases/phase-4.md +128 -0
- package/src/agents/specific_instructions/mlops_engineer/phases/phase-5.md +106 -0
- package/src/agents/specific_instructions/mlops_engineer/phases/phase-6.md +128 -0
- package/src/agents/specific_instructions/mlops_engineer/phases/phase-7.md +144 -0
- package/src/agents/specific_instructions/mlops_engineer/phases.md +671 -0
- package/src/agents/specific_instructions/mlops_engineer/review.md +164 -0
- package/src/agents/specific_instructions/mlops_engineer/service_mode.md +81 -0
- package/src/agents/specific_instructions/mlops_engineer/validation_checklist.md +151 -0
- package/src/agents/specific_instructions/researcher/critical_review.md +292 -0
- package/src/agents/specific_instructions/researcher/review_checklist.md +67 -0
- package/src/agents/specific_instructions/researcher/service_mode.md +224 -0
- package/src/agents/specific_instructions/shared/auto_verify_mode.md +141 -0
- package/src/agents/specific_instructions/shared/autonomous_research.md +1289 -0
- package/src/agents/specific_instructions/shared/behavioral_rules.md +36 -0
- package/src/agents/specific_instructions/shared/diverge_protocol.md +387 -0
- package/src/agents/specific_instructions/shared/engineering_guidelines.md +136 -0
- package/src/agents/specific_instructions/shared/experiment_versioning.md +184 -0
- package/src/agents/specific_instructions/shared/goal_mode.md +187 -0
- package/src/agents/specific_instructions/shared/incremental_testing.md +139 -0
- package/src/agents/specific_instructions/shared/intent_discovery.md +223 -0
- package/src/agents/specific_instructions/shared/join_path_protocol.md +168 -0
- package/src/agents/specific_instructions/shared/knowledge_checkpoint.md +83 -0
- package/src/agents/specific_instructions/shared/knowledge_harvest.md +220 -0
- package/src/agents/specific_instructions/shared/knowledge_retrieval.md +100 -0
- package/src/agents/specific_instructions/shared/notebook_walkthrough_protocol.md +367 -0
- package/src/agents/specific_instructions/shared/reviewer_verdict_protocol.md +74 -0
- package/src/agents/specific_instructions/shared/swarm_protocol.md +97 -0
- package/src/agents/specific_instructions/shared/validation_protocol.md +139 -0
- package/src/agents/specific_instructions/syn/arbiter.md +140 -0
- package/src/agents/specific_instructions/syn/brainstorm.md +550 -0
- package/src/agents/specific_instructions/syn/code_review.md +232 -0
- package/src/agents/specific_instructions/syn/diff.md +239 -0
- package/src/agents/specific_instructions/syn/final_review.md +65 -0
- package/src/agents/specific_instructions/syn/fixer.md +240 -0
- package/src/agents/specific_instructions/syn/free_form.md +130 -0
- package/src/agents/specific_instructions/syn/knowledge.md +468 -0
- package/src/agents/specific_instructions/syn/notebook_walkthrough.md +78 -0
- package/src/agents/specific_instructions/syn/panel_review.md +634 -0
- package/src/agents/specific_instructions/syn/pm.md +453 -0
- package/src/agents/specific_instructions/syn/pr_review.md +255 -0
- package/src/agents/specific_instructions/syn/slides.md +417 -0
- package/src/agents/syn.md +729 -0
- package/src/commands/academic.md +41 -0
- package/src/commands/ai-engineer.md +45 -0
- package/src/commands/analytics-engineer.md +48 -0
- package/src/commands/applied-ml-scientist.md +45 -0
- package/src/commands/backend-engineer.md +35 -0
- package/src/commands/bi-engineer.md +40 -0
- package/src/commands/brainstorm.md +24 -0
- package/src/commands/data-analyst.md +38 -0
- package/src/commands/data-engineer.md +37 -0
- package/src/commands/data-modeller.md +38 -0
- package/src/commands/data-scientist.md +38 -0
- package/src/commands/deep-learning-engineer.md +47 -0
- package/src/commands/end.md +49 -0
- package/src/commands/knowledge.md +24 -0
- package/src/commands/ml-engineer.md +42 -0
- package/src/commands/mlops-engineer.md +47 -0
- package/src/commands/notebook-walkthrough.md +58 -0
- package/src/commands/researcher.md +40 -0
- package/src/commands/resume.md +57 -0
- package/src/commands/review-pr.md +26 -0
- package/src/commands/shards-guide.md +41 -0
- package/src/commands/shards-ui.md +32 -0
- package/src/commands/shards.md +41 -0
- package/src/docs/01-getting-started/concepts.md +109 -0
- package/src/docs/01-getting-started/first-session.md +79 -0
- package/src/docs/01-getting-started/install.md +61 -0
- package/src/docs/02-agents/academic.md +71 -0
- package/src/docs/02-agents/ai-engineer.md +78 -0
- package/src/docs/02-agents/analytics-engineer.md +58 -0
- package/src/docs/02-agents/applied-ml-scientist.md +59 -0
- package/src/docs/02-agents/backend-engineer.md +58 -0
- package/src/docs/02-agents/bi-engineer.md +65 -0
- package/src/docs/02-agents/data-analyst.md +67 -0
- package/src/docs/02-agents/data-engineer.md +57 -0
- package/src/docs/02-agents/data-modeller.md +51 -0
- package/src/docs/02-agents/data-scientist.md +78 -0
- package/src/docs/02-agents/deep-learning-engineer.md +64 -0
- package/src/docs/02-agents/ml-engineer.md +80 -0
- package/src/docs/02-agents/mlops-engineer.md +59 -0
- package/src/docs/02-agents/overview.md +62 -0
- package/src/docs/02-agents/researcher.md +73 -0
- package/src/docs/02-agents/syn.md +88 -0
- package/src/docs/03-protocols/auto-verify.md +82 -0
- package/src/docs/03-protocols/autonomous-research.md +59 -0
- package/src/docs/03-protocols/behavioral-rules.md +35 -0
- package/src/docs/03-protocols/diverge.md +50 -0
- package/src/docs/03-protocols/engineering-guidelines.md +56 -0
- package/src/docs/03-protocols/experiment-versioning.md +38 -0
- package/src/docs/03-protocols/gate-pattern.md +65 -0
- package/src/docs/03-protocols/incremental-testing.md +68 -0
- package/src/docs/03-protocols/join-path.md +46 -0
- package/src/docs/03-protocols/knowledge-ledger.md +70 -0
- package/src/docs/03-protocols/reviewer-verdicts.md +39 -0
- package/src/docs/03-protocols/swarm.md +40 -0
- package/src/docs/03-protocols/validation.md +174 -0
- package/src/docs/04-ui/activity-bar.md +70 -0
- package/src/docs/04-ui/chat-pane.md +80 -0
- package/src/docs/04-ui/code-intel.md +62 -0
- package/src/docs/04-ui/file-editing.md +61 -0
- package/src/docs/04-ui/git.md +54 -0
- package/src/docs/04-ui/keybindings.md +79 -0
- package/src/docs/04-ui/knowledge-map.md +76 -0
- package/src/docs/04-ui/overview.md +93 -0
- package/src/docs/04-ui/panels.md +49 -0
- package/src/docs/04-ui/pinboard-selection.md +66 -0
- package/src/docs/04-ui/quick-open-palette.md +56 -0
- package/src/docs/04-ui/sessions.md +81 -0
- package/src/docs/04-ui/settings-permissions.md +56 -0
- package/src/docs/05-commands/reference.md +59 -0
- package/src/docs/06-outputs/directory-map.md +116 -0
- package/src/docs/07-workflows/ai-eval-first.md +57 -0
- package/src/docs/07-workflows/deep-study-to-production.md +76 -0
- package/src/docs/07-workflows/diverge-exploration.md +77 -0
- package/src/docs/07-workflows/quick-analysis.md +45 -0
- package/src/docs/08-integrations/claude-code-auto-mode.md +191 -0
- package/src/docs/08-integrations/google-slides.md +175 -0
- package/src/docs/README.md +30 -0
- package/src/docs/manifest.json +108 -0
- package/src/templates/analysis-template.md +20 -0
- package/src/templates/branch-report.md +46 -0
- package/src/templates/diff-report.md +88 -0
- package/src/templates/knowledge-index.md +7 -0
- package/src/templates/model-card-schema.json +186 -0
- package/src/templates/model-card-schema.md +88 -0
- package/src/templates/model-card.md +124 -0
- package/src/templates/project-plan.md +47 -0
- package/src/templates/project-specs.md +81 -0
- package/src/templates/report-template.md +43 -0
- package/src/templates/study-template.md +25 -0
- package/src/ui/cc-readonly.js +181 -0
- package/src/ui/chat-session.js +466 -0
- package/src/ui/css/base.css +136 -0
- package/src/ui/css/brainstorm.css +525 -0
- package/src/ui/css/chat.css +1405 -0
- package/src/ui/css/editor.css +546 -0
- package/src/ui/css/eval-dashboard.css +157 -0
- package/src/ui/css/experiment.css +237 -0
- package/src/ui/css/guide.css +186 -0
- package/src/ui/css/knowledge-map.css +383 -0
- package/src/ui/css/layout.css +431 -0
- package/src/ui/css/model-card.css +161 -0
- package/src/ui/css/notebook-walkthrough.css +271 -0
- package/src/ui/css/pr-review.css +403 -0
- package/src/ui/css/prompt-lab.css +325 -0
- package/src/ui/css/sessions.css +258 -0
- package/src/ui/css/sidebar.css +661 -0
- package/src/ui/css/terminal.css +113 -0
- package/src/ui/css/theme-light.css +542 -0
- package/src/ui/index.html +389 -0
- package/src/ui/js/agents.js +32 -0
- package/src/ui/js/bookmarks.js +230 -0
- package/src/ui/js/chat.js +1776 -0
- package/src/ui/js/code-intel.js +328 -0
- package/src/ui/js/command-palette.js +142 -0
- package/src/ui/js/events.js +591 -0
- package/src/ui/js/explorer.js +317 -0
- package/src/ui/js/file-view.js +477 -0
- package/src/ui/js/git.js +536 -0
- package/src/ui/js/guide.js +198 -0
- package/src/ui/js/hud.js +75 -0
- package/src/ui/js/init.js +351 -0
- package/src/ui/js/knowledge-map.js +906 -0
- package/src/ui/js/markdown.js +114 -0
- package/src/ui/js/monaco.js +164 -0
- package/src/ui/js/notebook-walkthrough.js +272 -0
- package/src/ui/js/notebook.js +448 -0
- package/src/ui/js/panels.js +2681 -0
- package/src/ui/js/pinboard.js +186 -0
- package/src/ui/js/quick-open.js +164 -0
- package/src/ui/js/selection-context.js +131 -0
- package/src/ui/js/sessions.js +256 -0
- package/src/ui/js/settings.js +476 -0
- package/src/ui/js/split-view.js +82 -0
- package/src/ui/js/state.js +343 -0
- package/src/ui/js/table.js +161 -0
- package/src/ui/js/tabs.js +284 -0
- package/src/ui/js/tabular.js +125 -0
- package/src/ui/js/terminal.js +354 -0
- package/src/ui/js/timeline.js +137 -0
- package/src/ui/js/utils.js +293 -0
- package/src/ui/notebook-kernel.py +790 -0
- package/src/ui/open-browser.js +55 -0
- package/src/ui/permission-pattern.js +42 -0
- package/src/ui/relay.js +513 -0
- package/src/ui/server.js +3072 -0
- package/src/ui/session-index.js +225 -0
- package/src/ui/shards_icon.png +0 -0
- package/src/ui/spawn-server.js +41 -0
- package/src/ui/symbol-index.js +813 -0
- package/src/ui/ui-push.js +177 -0
- package/tools/gate-hook/VALIDATION_SPEC.md +273 -0
- package/tools/gate-hook/__tests__/auto-verify.test.js +343 -0
- package/tools/gate-hook/auto-allowlist.js +179 -0
- package/tools/gate-hook/auto-state.js +68 -0
- package/tools/gate-hook/classify.js +21 -0
- package/tools/gate-hook/log.js +57 -0
- package/tools/gate-hook/parser.js +205 -0
- package/tools/gate-hook/sql-guard.js +230 -0
- package/tools/gate-hook/state.js +170 -0
- package/tools/gate-hook/sweep.js +139 -0
- package/tools/gate-hook/transcript.js +45 -0
- package/tools/gate-hook/validation.js +321 -0
- package/tools/gate-hook.js +475 -0
- package/tools/install.js +914 -0
- package/tools/shards-gates.js +311 -0
- package/tools/shards-sessions.js +261 -0
- package/tools/shards-ui.js +377 -0
|
@@ -0,0 +1,98 @@
|
|
|
1
|
+
> **Previous:** phase-1.md confirmed
|
|
2
|
+
> **Next:** phase-3.md (read only after this phase's gate is confirmed)
|
|
3
|
+
|
|
4
|
+
---
|
|
5
|
+
|
|
6
|
+
## Create Mode — Phase 2: Training Protocol (Gated)
|
|
7
|
+
|
|
8
|
+
Goal: Fully specify training before writing a line of code.
|
|
9
|
+
|
|
10
|
+
Define:
|
|
11
|
+
|
|
12
|
+
1. **Loss function:** Formula + justification. Why this loss for this task?
|
|
13
|
+
Known failure modes?
|
|
14
|
+
|
|
15
|
+
2. **Optimizer:** AdamW is the default. State the reason explicitly if
|
|
16
|
+
deviating. Include:
|
|
17
|
+
- Weight decay value and rationale (decoupled from LR per Loshchilov &
|
|
18
|
+
Hutter, 2019)
|
|
19
|
+
- β₁, β₂, ε values if non-default, with rationale
|
|
20
|
+
|
|
21
|
+
3. **Learning rate schedule:**
|
|
22
|
+
- Warmup: number of steps and rationale
|
|
23
|
+
- Decay strategy: cosine, linear, polynomial — with rationale
|
|
24
|
+
- Peak LR: concrete value with concrete justification (not "tune it")
|
|
25
|
+
- Minimum LR (if applicable)
|
|
26
|
+
|
|
27
|
+
4. **Regularization:**
|
|
28
|
+
- Dropout: rate and placement (attention dropout vs residual dropout vs
|
|
29
|
+
classifier dropout — these are different)
|
|
30
|
+
- Label smoothing (if classification): value and rationale
|
|
31
|
+
- Stochastic depth (if applicable): survival probability
|
|
32
|
+
- Weight decay already specified in optimizer
|
|
33
|
+
|
|
34
|
+
5. **Data augmentation table:**
|
|
35
|
+
|
|
36
|
+
| Transform | Parameters | Invariance Encoded | Apply to Val? |
|
|
37
|
+
|-----------|-----------|-------------------|---------------|
|
|
38
|
+
| <name> | <params> | <what it teaches> | <yes | no> |
|
|
39
|
+
|
|
40
|
+
6. **Batch configuration:**
|
|
41
|
+
- Effective batch size (global)
|
|
42
|
+
- Per-GPU batch size
|
|
43
|
+
- Gradient accumulation steps (if VRAM-constrained)
|
|
44
|
+
|
|
45
|
+
7. **Checkpoint strategy:** best validation metric, every N epochs, or both.
|
|
46
|
+
|
|
47
|
+
### Document Phase 2
|
|
48
|
+
|
|
49
|
+
Append to `project-specs.md`:
|
|
50
|
+
|
|
51
|
+
```markdown
|
|
52
|
+
## Phase 2: Training Protocol
|
|
53
|
+
|
|
54
|
+
### Loss Function
|
|
55
|
+
- **Formula:** <L = ...>
|
|
56
|
+
- **Justification:** <why this loss for this task>
|
|
57
|
+
- **Known failure modes:** <class imbalance, optimization landscape issues, etc.>
|
|
58
|
+
|
|
59
|
+
### Optimizer
|
|
60
|
+
- **Optimizer:** <AdamW | other>
|
|
61
|
+
- **Deviation rationale:** <if not AdamW, why>
|
|
62
|
+
- **Weight decay:** <value> — <rationale>
|
|
63
|
+
- **β₁, β₂, ε:** <values if non-default>
|
|
64
|
+
|
|
65
|
+
### Learning Rate Schedule
|
|
66
|
+
- **Warmup:** <N steps> — <rationale>
|
|
67
|
+
- **Decay:** <cosine | linear | polynomial> — <rationale>
|
|
68
|
+
- **Peak LR:** <value> — <justification>
|
|
69
|
+
- **Minimum LR:** <value or "none">
|
|
70
|
+
|
|
71
|
+
### Regularization
|
|
72
|
+
- **Dropout:** rate=<X>, placement=<where>
|
|
73
|
+
- **Label smoothing:** <value | N/A>
|
|
74
|
+
- **Stochastic depth:** survival_prob=<X | N/A>
|
|
75
|
+
|
|
76
|
+
### Augmentation
|
|
77
|
+
| Transform | Parameters | Invariance | Val? |
|
|
78
|
+
|-----------|-----------|-----------|------|
|
|
79
|
+
| <name> | <params> | <invariance> | <yes/no> |
|
|
80
|
+
|
|
81
|
+
### Batch Configuration
|
|
82
|
+
- **Effective batch size:** <N>
|
|
83
|
+
- **Per-GPU batch size:** <N>
|
|
84
|
+
- **Gradient accumulation:** <N steps | none>
|
|
85
|
+
|
|
86
|
+
### Checkpoint Strategy
|
|
87
|
+
<best val metric | every N epochs | both — rationale>
|
|
88
|
+
```
|
|
89
|
+
|
|
90
|
+
::GATE:: id=deep-learning-engineer-phase-2 phase=2 kind=phase
|
|
91
|
+
Read this section back to the user. Stop here — do not begin the next phase or output any further content. Wait for the user to explicitly confirm before proceeding. Do not interpret silence or partial agreement as confirmation.
|
|
92
|
+
::ENDGATE::
|
|
93
|
+
|
|
94
|
+
---
|
|
95
|
+
|
|
96
|
+
## When this gate is confirmed
|
|
97
|
+
|
|
98
|
+
Read `.claude/agents/specific_instructions/deep_learning_engineer/phases/phase-3.md` in full and follow its instructions starting from Phase 3. Do not pre-read further phase files.
|
|
@@ -0,0 +1,76 @@
|
|
|
1
|
+
> **Previous:** phase-2.md confirmed
|
|
2
|
+
> **Next:** phase-4.md (read only after this phase's gate is confirmed)
|
|
3
|
+
|
|
4
|
+
---
|
|
5
|
+
|
|
6
|
+
## Create Mode — Phase 3: Implementation Specification (Gated)
|
|
7
|
+
|
|
8
|
+
Goal: Translate architecture and training protocol into an engineering plan.
|
|
9
|
+
|
|
10
|
+
Define:
|
|
11
|
+
|
|
12
|
+
1. **Framework:** PyTorch vs JAX with rationale. Not a preference — a reason
|
|
13
|
+
tied to the training procedure (custom CUDA, vmap, multi-host TPU, etc.).
|
|
14
|
+
|
|
15
|
+
2. **Code structure:** What each `src/` file contains, `forward()` signature
|
|
16
|
+
with input and output shapes, module interfaces.
|
|
17
|
+
|
|
18
|
+
3. **Hardware config:**
|
|
19
|
+
- fp16 vs bf16: bf16 preferred for most modern GPUs (A100, H100, RTX 30xx+);
|
|
20
|
+
fp16 for older hardware — state the reason
|
|
21
|
+
- Gradient checkpointing: yes if VRAM is constrained; quantify the compute
|
|
22
|
+
overhead
|
|
23
|
+
- `torch.compile`: compatible with the architecture? Expected speedup?
|
|
24
|
+
|
|
25
|
+
4. **Experiment tracking:** tool (wandb / tensorboard / MLflow), key metrics
|
|
26
|
+
logged per step and per epoch, visualization plan.
|
|
27
|
+
|
|
28
|
+
5. **Inference plan:** serving format, quantization (int8 / int4 / GPTQ if
|
|
29
|
+
applicable), expected latency delta from quantization.
|
|
30
|
+
|
|
31
|
+
### Document Phase 3
|
|
32
|
+
|
|
33
|
+
Append to `project-specs.md`:
|
|
34
|
+
|
|
35
|
+
```markdown
|
|
36
|
+
## Phase 3: Implementation Specification
|
|
37
|
+
|
|
38
|
+
### Framework
|
|
39
|
+
- **Framework:** <PyTorch | JAX>
|
|
40
|
+
- **Rationale:** <concrete reason tied to training procedure>
|
|
41
|
+
|
|
42
|
+
### Code Structure
|
|
43
|
+
- **model.py:** <what it contains, forward() signature with shapes>
|
|
44
|
+
- **dataset.py:** <Dataset class, transforms, get_dataloaders() fn>
|
|
45
|
+
- **train.py:** <training loop, optimizer/scheduler construction, checkpoint logic>
|
|
46
|
+
- **evaluate.py:** <evaluation loop, metric computation, predict() fn>
|
|
47
|
+
- **configs/config.yaml:** <all hyperparameters from Phases 1-2>
|
|
48
|
+
|
|
49
|
+
### Hardware Config
|
|
50
|
+
- **Precision:** <bf16 | fp16> — <rationale>
|
|
51
|
+
- **Gradient checkpointing:** <yes — ~X% compute overhead | no>
|
|
52
|
+
- **torch.compile:** <yes — expected ~X% speedup | no — incompatible because>
|
|
53
|
+
|
|
54
|
+
### Experiment Tracking
|
|
55
|
+
- **Tool:** <wandb | tensorboard | MLflow>
|
|
56
|
+
- **Metrics per step:** <loss, grad norm, LR>
|
|
57
|
+
- **Metrics per epoch:** <val loss, primary metric, secondary metrics>
|
|
58
|
+
- **Visualizations:** <loss curves, confusion matrix, attention maps, etc.>
|
|
59
|
+
|
|
60
|
+
### Inference Plan
|
|
61
|
+
- **Serving format:** <ONNX | TorchScript | TensorRT | HuggingFace | none>
|
|
62
|
+
- **Quantization:** <int8 | int4 | none> — <latency delta estimate>
|
|
63
|
+
- **Expected inference latency:** ~<X>ms per sample on <hardware>
|
|
64
|
+
```
|
|
65
|
+
|
|
66
|
+
**DIVERGE check:** If you identified 2-3 mutually exclusive neural architectures (e.g., different backbone families, fundamentally different training paradigms) that are genuinely equally viable, you MAY propose a DIVERGE fork. Read `.claude/agents/specific_instructions/shared/diverge_protocol.md` and follow its DIVERGE Proposal Gate. If confirmed, branches execute autonomously through the remaining phases. After convergence and promotion, resume at Phase 4. If declined or not applicable, continue normally.
|
|
67
|
+
|
|
68
|
+
::GATE:: id=deep-learning-engineer-phase-3 phase=3 kind=phase
|
|
69
|
+
Read this section back to the user. Stop here — do not begin the next phase or output any further content. Wait for the user to explicitly confirm before proceeding. Do not interpret silence or partial agreement as confirmation.
|
|
70
|
+
::ENDGATE::
|
|
71
|
+
|
|
72
|
+
---
|
|
73
|
+
|
|
74
|
+
## When this gate is confirmed
|
|
75
|
+
|
|
76
|
+
Read `.claude/agents/specific_instructions/deep_learning_engineer/phases/phase-4.md` in full and follow its instructions starting from Phase 4. Do not pre-read further phase files.
|
|
@@ -0,0 +1,128 @@
|
|
|
1
|
+
> **Previous:** phase-3.md confirmed
|
|
2
|
+
> **Next:** phase-5.md (read only after this phase's gate is confirmed)
|
|
3
|
+
|
|
4
|
+
---
|
|
5
|
+
|
|
6
|
+
## Create Mode — Phase 4: Execute (Gated)
|
|
7
|
+
|
|
8
|
+
Goal: Build the model implementation.
|
|
9
|
+
|
|
10
|
+
**Context checkpoint:** Before building, prompt the user:
|
|
11
|
+
|
|
12
|
+
"Planning's locked — good moment to run `/compact` or `/clear` before we start
|
|
13
|
+
executing. I'll be working from project-specs.md from here. Say the word when
|
|
14
|
+
you're ready."
|
|
15
|
+
|
|
16
|
+
Wait for any signal from the user before beginning build steps.
|
|
17
|
+
|
|
18
|
+
**Knowledge re-check:** Follow `.claude/agents/specific_instructions/shared/knowledge_checkpoint.md` before building.
|
|
19
|
+
|
|
20
|
+
### Incremental testing — checkpoint gates between components
|
|
21
|
+
|
|
22
|
+
Follow `.claude/agents/specific_instructions/shared/incremental_testing.md` during this build. Each component below is a checkpoint seam — after you write and execute a component, emit a `kind=checkpoint` gate fence (template below) and wait for user confirmation before starting the next component. Do not leave run-all until the end: test each component in isolation as you build it.
|
|
23
|
+
|
|
24
|
+
Checkpoint gate fence — emit exactly this shape. Both `::GATE::` and `::ENDGATE::` fences are required, as are all three attributes (`id`, `phase`, `kind`). No prose outside the fence.
|
|
25
|
+
|
|
26
|
+
```
|
|
27
|
+
::GATE:: id=<agent-name>-phase-<N>-checkpoint-<component> phase=<N> kind=checkpoint
|
|
28
|
+
Component: <human-readable name>
|
|
29
|
+
Test command: <exact command you ran>
|
|
30
|
+
Evidence:
|
|
31
|
+
- <measured fact 1, e.g. "df.shape = (48211, 47)">
|
|
32
|
+
- <measured fact 2, e.g. "null rate on join key = 0.00%">
|
|
33
|
+
- <measured fact 3, e.g. "sample head matches expected schema">
|
|
34
|
+
Status: PASS | FAIL — <one-line summary>
|
|
35
|
+
Next: <what you'll build after this is confirmed>
|
|
36
|
+
Stop here — await explicit confirmation before writing the next component.
|
|
37
|
+
::ENDGATE::
|
|
38
|
+
```
|
|
39
|
+
|
|
40
|
+
Expected checkpoint gate IDs for this phase (emit in order as you build):
|
|
41
|
+
|
|
42
|
+
- `deep-learning-engineer-phase-4-checkpoint-data` — dataset load + sample batch visualization; shape assertions hold; class distribution (if classification) matches prior.
|
|
43
|
+
- `deep-learning-engineer-phase-4-checkpoint-forward` — model forward pass on a dummy batch; shapes at each component match expectation; no NaNs.
|
|
44
|
+
- `deep-learning-engineer-phase-4-checkpoint-smoke-train` — smoke-fit on ≤1% of data for a handful of epochs; loss decreases, gradient norms finite; tiny-batch overfitting sanity check passes.
|
|
45
|
+
- `deep-learning-engineer-phase-4-checkpoint-full-train` — full training completes on target hardware; peak memory within budget; best validation metric logged.
|
|
46
|
+
- `deep-learning-engineer-phase-4-checkpoint-eval` — evaluation + diagnostics (gradient history, activation stats, dead-neuron check) produced and reviewed.
|
|
47
|
+
|
|
48
|
+
The hook blocks all non-read tools while a checkpoint is open. If a checkpoint fails, diagnose and re-emit with updated evidence before advancing. Use the fence body format shown above (Component / Test command / Evidence / Status / Next).
|
|
49
|
+
|
|
50
|
+
Then create:
|
|
51
|
+
|
|
52
|
+
**`models/<project_name>/notebooks/model_development.ipynb`**
|
|
53
|
+
|
|
54
|
+
Structure:
|
|
55
|
+
1. **Setup** — imports, config load, device setup, seed setting
|
|
56
|
+
2. **Dataset Sanity Check** — data loading, shape assertions, sample batch
|
|
57
|
+
visualization, class distribution (if classification)
|
|
58
|
+
3. **Model Trace** — instantiate model, run forward pass with dummy input,
|
|
59
|
+
print shape at each major component via hooks or explicit prints
|
|
60
|
+
4. **Training Loop** — full training with logging: loss, eval metric, gradient
|
|
61
|
+
norm, LR per epoch; inline loss curves after training
|
|
62
|
+
5. **Evaluation** — metric table vs baseline, confusion matrix or equivalent,
|
|
63
|
+
per-class breakdown if applicable
|
|
64
|
+
6. **Diagnostics** — gradient norm history plot, activation statistics,
|
|
65
|
+
loss curve analysis, dead neuron check (if ReLU backbone)
|
|
66
|
+
|
|
67
|
+
**`models/<project_name>/src/model.py`**
|
|
68
|
+
- Every `forward()` method has a shape comment on the return tensor
|
|
69
|
+
- No magic numbers — all sizes come from config
|
|
70
|
+
- Normalization and dropout layers instantiated in `__init__`, applied in `forward()`
|
|
71
|
+
|
|
72
|
+
**`models/<project_name>/src/dataset.py`**
|
|
73
|
+
- `Dataset` class with `__len__` and `__getitem__`
|
|
74
|
+
- Transform pipeline built from Phase 2 augmentation table
|
|
75
|
+
- `get_dataloaders(config)` convenience function
|
|
76
|
+
|
|
77
|
+
**`models/<project_name>/src/train.py`**
|
|
78
|
+
- `train_epoch(model, loader, optimizer, scheduler, device)` function
|
|
79
|
+
- Optimizer and scheduler construction
|
|
80
|
+
- Checkpoint logic: save best validation metric, load from checkpoint
|
|
81
|
+
|
|
82
|
+
**`models/<project_name>/src/evaluate.py`**
|
|
83
|
+
- `evaluate(model, loader, device)` evaluation loop
|
|
84
|
+
- Primary and secondary metric computation
|
|
85
|
+
- `predict(model, sample, device)` single-sample inference function
|
|
86
|
+
|
|
87
|
+
**`models/<project_name>/configs/config.yaml`**
|
|
88
|
+
- All hyperparameters from Phases 1, 2, and 3
|
|
89
|
+
- No hardcoded values in src/ — everything references config
|
|
90
|
+
|
|
91
|
+
**`models/<project_name>/requirements.txt`**
|
|
92
|
+
- Pinned major dependencies (torch==X.Y, torchvision==X.Y, etc.)
|
|
93
|
+
|
|
94
|
+
### Document Phase 4
|
|
95
|
+
|
|
96
|
+
Append to `project-specs.md`:
|
|
97
|
+
|
|
98
|
+
```markdown
|
|
99
|
+
## Phase 4: Build Log
|
|
100
|
+
|
|
101
|
+
- **Notebook:** `notebooks/model_development.ipynb`
|
|
102
|
+
- **Source modules:** model.py, dataset.py, train.py, evaluate.py
|
|
103
|
+
- **Training run summary:**
|
|
104
|
+
- Hardware: <GPU model, VRAM>
|
|
105
|
+
- Precision: <bf16 | fp16 | fp32>
|
|
106
|
+
- Effective batch size: <N>
|
|
107
|
+
- Steps / epochs: <N>
|
|
108
|
+
- Peak GPU memory: ~<X>GB
|
|
109
|
+
- Loss trajectory: <converged at epoch N | diverged | oscillating — describe>
|
|
110
|
+
- Best validation metric: <metric name>=<value> at epoch <N>
|
|
111
|
+
- **Baseline comparison:**
|
|
112
|
+
| Model | <Metric> | Notes |
|
|
113
|
+
|-------|---------|-------|
|
|
114
|
+
| Baseline (<type>) | <value> | — |
|
|
115
|
+
| **Ours** | <value> | — |
|
|
116
|
+
- **Gradient diagnostics:** <healthy convergence | anomalies observed — describe>
|
|
117
|
+
- **Known implementation limitations:** <what prototype doesn't handle>
|
|
118
|
+
```
|
|
119
|
+
|
|
120
|
+
::GATE:: id=deep-learning-engineer-phase-4 phase=4 kind=phase validates=deep_learning_engineer
|
|
121
|
+
Read this section back to the user. Stop here — do not begin the next phase or output any further content. Wait for the user to explicitly confirm before proceeding. Do not interpret silence or partial agreement as confirmation.
|
|
122
|
+
::ENDGATE::
|
|
123
|
+
|
|
124
|
+
---
|
|
125
|
+
|
|
126
|
+
## When this gate is confirmed
|
|
127
|
+
|
|
128
|
+
Read `.claude/agents/specific_instructions/deep_learning_engineer/phases/phase-5.md` in full and follow its instructions starting from Phase 5. Do not pre-read further phase files.
|
|
@@ -0,0 +1,292 @@
|
|
|
1
|
+
> **Previous:** phase-4.md confirmed
|
|
2
|
+
> **Next:** This is the final phase — follow the Syn sign-off instructions in this phase to close the project.
|
|
3
|
+
|
|
4
|
+
---
|
|
5
|
+
|
|
6
|
+
## Create Mode — Phase 5: Review and Handoff (Gated)
|
|
7
|
+
|
|
8
|
+
Goal: Dual specialist review before writing the final report.
|
|
9
|
+
|
|
10
|
+
Both reviews happen before writing `report.md`. If either raises a blocking
|
|
11
|
+
concern, discuss with the user and revise before proceeding.
|
|
12
|
+
|
|
13
|
+
**ML Engineer review (production infrastructure and deployment feasibility):**
|
|
14
|
+
|
|
15
|
+
Tell the user: "Requesting ML Engineer review for production infrastructure
|
|
16
|
+
and deployment feasibility..."
|
|
17
|
+
|
|
18
|
+
```
|
|
19
|
+
Task(
|
|
20
|
+
subagent_type="ml-engineer",
|
|
21
|
+
description="Production readiness review for deep learning model: <project name>",
|
|
22
|
+
prompt="I am the Deep Learning Engineer shard. I have built a custom deep
|
|
23
|
+
learning model and need a production infrastructure review.
|
|
24
|
+
|
|
25
|
+
Project: <project name>
|
|
26
|
+
Directory: models/<project_name>/
|
|
27
|
+
Specs: models/<project_name>/project-specs.md
|
|
28
|
+
|
|
29
|
+
Summary:
|
|
30
|
+
- Task: <input/output tensor shapes from Phase 0>
|
|
31
|
+
- Architecture: <selected backbone + head from Phase 1>
|
|
32
|
+
- Parameters: ~<N>M
|
|
33
|
+
- Training: <optimizer, LR schedule, augmentation summary from Phase 2>
|
|
34
|
+
- Hardware: <GPU, precision, gradient checkpointing from Phase 3>
|
|
35
|
+
- Results: <best validation metric vs baseline from Phase 4>
|
|
36
|
+
- Serving format: <from Phase 3>
|
|
37
|
+
|
|
38
|
+
Please review for production readiness:
|
|
39
|
+
1. Is the serving format (or absence of one) appropriate for the stated
|
|
40
|
+
latency budget?
|
|
41
|
+
2. Are there infrastructure or pipeline concerns for integrating this model
|
|
42
|
+
into production?
|
|
43
|
+
3. Is the model size and inference cost acceptable for the stated hardware
|
|
44
|
+
constraints?
|
|
45
|
+
4. What monitoring and retraining triggers would you recommend?
|
|
46
|
+
5. Are there production failure modes (data drift, distribution shift,
|
|
47
|
+
cold start) not addressed in the current design?
|
|
48
|
+
|
|
49
|
+
Please read project-specs.md for full context."
|
|
50
|
+
)
|
|
51
|
+
```
|
|
52
|
+
|
|
53
|
+
**Applied ML Scientist review (methodology soundness and cutting-edge alternatives):**
|
|
54
|
+
|
|
55
|
+
Tell the user: "Requesting Applied ML Scientist review for methodology soundness
|
|
56
|
+
and cutting-edge alternatives assessment..."
|
|
57
|
+
|
|
58
|
+
```
|
|
59
|
+
Task(
|
|
60
|
+
subagent_type="applied-ml-scientist",
|
|
61
|
+
description="Methodology review for deep learning model: <project name>",
|
|
62
|
+
prompt="I am the Deep Learning Engineer shard. I have built a custom deep
|
|
63
|
+
learning model and need a methodology and theory review.
|
|
64
|
+
|
|
65
|
+
Project: <project name>
|
|
66
|
+
Directory: models/<project_name>/
|
|
67
|
+
|
|
68
|
+
Summary:
|
|
69
|
+
- Task: <problem description and input/output from Phase 0>
|
|
70
|
+
- Core hypothesis: <why this architecture for this data>
|
|
71
|
+
- Architecture: <selected approach from Phase 1>
|
|
72
|
+
- Loss function: <formula and justification from Phase 2>
|
|
73
|
+
- Results: <metric table from Phase 4>
|
|
74
|
+
|
|
75
|
+
Please review for methodology soundness:
|
|
76
|
+
1. Is the inductive bias argument for the selected architecture sound given
|
|
77
|
+
the data structure?
|
|
78
|
+
2. Are there recent methods (post-2022) that would clearly outperform this
|
|
79
|
+
approach for this problem type?
|
|
80
|
+
3. Is the loss function well-aligned with the task objective?
|
|
81
|
+
4. Are there theoretical gaps in the training protocol (optimizer choice,
|
|
82
|
+
regularization, augmentation strategy)?
|
|
83
|
+
5. What experiments would most efficiently validate or invalidate the core
|
|
84
|
+
design hypothesis?
|
|
85
|
+
|
|
86
|
+
Please read models/<project_name>/project-specs.md for full context."
|
|
87
|
+
)
|
|
88
|
+
```
|
|
89
|
+
|
|
90
|
+
**MLOps Engineer review (production platform and operationalization):**
|
|
91
|
+
|
|
92
|
+
Tell the user: "And finally, asking the MLOps Engineer to review the production
|
|
93
|
+
platform requirements and operationalization plan..."
|
|
94
|
+
|
|
95
|
+
```
|
|
96
|
+
Task(
|
|
97
|
+
subagent_type="mlops-engineer",
|
|
98
|
+
description="Production platform review for deep learning model: <project name>",
|
|
99
|
+
prompt="I am the Deep Learning Engineer shard. I have built a custom deep
|
|
100
|
+
learning model and need a production platform and operationalization review.
|
|
101
|
+
|
|
102
|
+
Project: <project name>
|
|
103
|
+
Directory: models/<project_name>/
|
|
104
|
+
Specs: models/<project_name>/project-specs.md
|
|
105
|
+
|
|
106
|
+
Summary:
|
|
107
|
+
- Architecture: <selected backbone + head from Phase 1>
|
|
108
|
+
- Parameters: ~<N>M
|
|
109
|
+
- Serving format: <from Phase 3>
|
|
110
|
+
- Hardware: <GPU, precision from Phase 3>
|
|
111
|
+
- Results: <best validation metric vs baseline from Phase 4>
|
|
112
|
+
|
|
113
|
+
Please review:
|
|
114
|
+
1. Is the serving format appropriate for the stated latency budget and
|
|
115
|
+
operational constraints?
|
|
116
|
+
2. What CI/CD pipeline would you recommend for retraining and model
|
|
117
|
+
registry management?
|
|
118
|
+
3. Is the experiment tracking and model versioning plan sufficient for
|
|
119
|
+
production operation?
|
|
120
|
+
4. What monitoring and retraining triggers would you build for this model?
|
|
121
|
+
5. What infrastructure is needed that isn't yet in the plan?
|
|
122
|
+
|
|
123
|
+
Please read project-specs.md for full context."
|
|
124
|
+
)
|
|
125
|
+
```
|
|
126
|
+
|
|
127
|
+
Append MLOps Engineer's review to specs.
|
|
128
|
+
|
|
129
|
+
**Backend Engineer code review (Python artifacts):**
|
|
130
|
+
|
|
131
|
+
Glob the project directory (`models/<project_name>/`) for `.py` and `.ipynb` files.
|
|
132
|
+
If any are found:
|
|
133
|
+
|
|
134
|
+
Tell the user: "Before we write the report, the Backend Engineer is reviewing
|
|
135
|
+
the Python artifacts. Code quality is not optional."
|
|
136
|
+
|
|
137
|
+
```
|
|
138
|
+
Task(
|
|
139
|
+
subagent_type="backend-engineer",
|
|
140
|
+
description="Python code review for [project_name]",
|
|
141
|
+
prompt="You are in SERVICE MODE. Review the following Python files in the
|
|
142
|
+
project at models/[project_name]/. Read project-specs.md first for context.
|
|
143
|
+
Files to review: [list of .py and .ipynb files found]"
|
|
144
|
+
)
|
|
145
|
+
```
|
|
146
|
+
|
|
147
|
+
Append the Backend Engineer's review to project-specs.md. If no Python files are
|
|
148
|
+
found, skip this step.
|
|
149
|
+
|
|
150
|
+
**Consult Syn for final sign-off:**
|
|
151
|
+
|
|
152
|
+
Tell the user: "I'm asking Syn to review the deep learning model design,
|
|
153
|
+
training protocol, and results before we close..."
|
|
154
|
+
|
|
155
|
+
```
|
|
156
|
+
Task(
|
|
157
|
+
subagent_type="syn",
|
|
158
|
+
description="Final review of deep learning model: <project name>",
|
|
159
|
+
prompt="I am the Deep Learning Engineer shard. I have completed a custom
|
|
160
|
+
deep learning model project. Please review and provide APPROVED / NEEDS
|
|
161
|
+
REVISION / BLOCKED.
|
|
162
|
+
|
|
163
|
+
Project: <project name>
|
|
164
|
+
Directory: models/<project_name>/
|
|
165
|
+
Specs: models/<project_name>/project-specs.md
|
|
166
|
+
|
|
167
|
+
Summary:
|
|
168
|
+
- Task: <input → output from Phase 0>
|
|
169
|
+
- Architecture: <selected backbone + head from Phase 1>, ~<N>M parameters
|
|
170
|
+
- Training: <optimizer, LR schedule, loss function from Phase 2>
|
|
171
|
+
- Hardware: <GPU, precision, gradient checkpointing from Phase 3>
|
|
172
|
+
- Results: <best validation metric vs baseline from Phase 4>
|
|
173
|
+
- Known limitations: <from Phase 4>
|
|
174
|
+
|
|
175
|
+
Reviewer verdicts already collected:
|
|
176
|
+
- ML Engineer (production readiness): <DEPLOY | OPTIMIZE | REDESIGN> — <one-line reason>
|
|
177
|
+
- Applied ML Scientist (methodology): <Sound | Consider Alternatives | Revise> — <one-line reason>
|
|
178
|
+
- MLOps Engineer (operationalization): <Approved | Concerns | Redesign needed> — <one-line reason>
|
|
179
|
+
- Backend Engineer (code quality): <Clean | Minor Issues | Refactor Required | Blocked | N/A> — <one-line reason>
|
|
180
|
+
|
|
181
|
+
Please read project-specs.md for full context and confirm whether the
|
|
182
|
+
project is ready to close given the reviewer verdicts above."
|
|
183
|
+
)
|
|
184
|
+
```
|
|
185
|
+
|
|
186
|
+
Append Syn's verdict to project-specs.md. If Syn returns NEEDS REVISION or
|
|
187
|
+
BLOCKED, discuss with the user and address before proceeding.
|
|
188
|
+
|
|
189
|
+
**Multi-reviewer conflict protocol:**
|
|
190
|
+
|
|
191
|
+
If no reviewer returns a blocking verdict (REDESIGN, Revise, or Redesign needed):
|
|
192
|
+
→ Proceed to report.md. Document all three verdicts.
|
|
193
|
+
|
|
194
|
+
If all reviewers with blocking concerns agree on the same root cause:
|
|
195
|
+
→ Discuss with the user and revise the binding issue before proceeding.
|
|
196
|
+
|
|
197
|
+
If reviewers disagree — one or two block while the other(s) do not:
|
|
198
|
+
→ Present the conflict explicitly to the user:
|
|
199
|
+
|
|
200
|
+
"The reviewers disagree:
|
|
201
|
+
- ML Engineer verdict: [DEPLOY | OPTIMIZE | REDESIGN] — [key reason]
|
|
202
|
+
- Applied ML Scientist verdict: [Sound | Consider Alternatives | Revise] — [key reason]
|
|
203
|
+
- MLOps Engineer verdict: [Approved | Concerns | Redesign needed] — [key reason]
|
|
204
|
+
|
|
205
|
+
This is a genuine constraint conflict. Which is the binding constraint for
|
|
206
|
+
this project: production feasibility, methodological rigor, or operational
|
|
207
|
+
readiness? Your answer determines what we fix first."
|
|
208
|
+
|
|
209
|
+
Document the user's decision in project-specs.md:
|
|
210
|
+
|
|
211
|
+
**Reviewer conflict resolution:** Production-first | Methodology-first | Operations-first | User override — <rationale>
|
|
212
|
+
|
|
213
|
+
Then address the binding constraint before proceeding. If the non-binding concern
|
|
214
|
+
remains unresolved after iteration, note it explicitly in report.md Limitations.
|
|
215
|
+
|
|
216
|
+
**Create `models/<project_name>/report.md`:**
|
|
217
|
+
|
|
218
|
+
```markdown
|
|
219
|
+
# <Project Name> — Deep Learning Engineering Report
|
|
220
|
+
|
|
221
|
+
## System Specification
|
|
222
|
+
- **Task:** <input → output>
|
|
223
|
+
- **Architecture:** <name, parameter count>
|
|
224
|
+
- **Training:** <optimizer, schedule, key regularization>
|
|
225
|
+
- **Hardware:** <GPU, precision>
|
|
226
|
+
|
|
227
|
+
## Architecture Rationale
|
|
228
|
+
<Inductive bias argument for the selected architecture. Cite the paper that
|
|
229
|
+
established why this architecture class works for this data modality.>
|
|
230
|
+
|
|
231
|
+
### Candidate Comparison
|
|
232
|
+
<Table from Phase 1>
|
|
233
|
+
|
|
234
|
+
## Training Protocol Rationale
|
|
235
|
+
<Justification for optimizer, LR schedule, loss function, and augmentation
|
|
236
|
+
choices. Cite papers where relevant (Loshchilov & Hutter, 2019 for decoupled
|
|
237
|
+
weight decay; Goyal et al., 2017 for LR scaling, etc.)>
|
|
238
|
+
|
|
239
|
+
## Results
|
|
240
|
+
<Metric table vs baseline from Phase 4>
|
|
241
|
+
<Training dynamics: loss trajectory, convergence epoch, gradient norm behavior>
|
|
242
|
+
<Ablation results if run>
|
|
243
|
+
|
|
244
|
+
## Production Readiness
|
|
245
|
+
**ML Engineer verdict:** <DEPLOY | OPTIMIZE | REDESIGN>
|
|
246
|
+
<Summary of infrastructure and deployment assessment>
|
|
247
|
+
<Required changes before production: ordered by priority>
|
|
248
|
+
|
|
249
|
+
## Methodology Review
|
|
250
|
+
**Applied ML Scientist verdict:** <Sound | Consider Alternatives | Revise>
|
|
251
|
+
<Summary of theoretical soundness assessment>
|
|
252
|
+
<Literature gaps or superior alternatives identified>
|
|
253
|
+
|
|
254
|
+
## Operations Review
|
|
255
|
+
**MLOps Engineer verdict:** <Approved | Concerns | Redesign needed>
|
|
256
|
+
<Summary of production platform and operationalization assessment>
|
|
257
|
+
<CI/CD, monitoring, and retraining gaps identified>
|
|
258
|
+
|
|
259
|
+
## Code Review
|
|
260
|
+
**Backend Engineer verdict:** <Clean | Minor Issues | Refactor Required | Blocked | N/A — no Python artifacts>
|
|
261
|
+
<Summary of code review findings, or "No Python artifacts found.">
|
|
262
|
+
|
|
263
|
+
## Limitations
|
|
264
|
+
<What the current implementation does not handle:>
|
|
265
|
+
- <Hardware assumption: trained on X, deployed assumptions unclear>
|
|
266
|
+
- <Data quality: assumes Y, not validated for Z>
|
|
267
|
+
- <Scale: prototype at N examples; behavior at 10× unknown>
|
|
268
|
+
|
|
269
|
+
## Next Steps
|
|
270
|
+
<Ordered by expected impact:>
|
|
271
|
+
1. <experiment or engineering step>
|
|
272
|
+
2. <experiment or engineering step>
|
|
273
|
+
3. <experiment or engineering step>
|
|
274
|
+
|
|
275
|
+
## Knowledge Harvested
|
|
276
|
+
- <title> → .shards/knowledge/<type>/<filename>.md
|
|
277
|
+
- Or: None — project did not produce reusable knowledge
|
|
278
|
+
```
|
|
279
|
+
|
|
280
|
+
**Knowledge harvest.** Before closing, extract reusable knowledge from this project.
|
|
281
|
+
Read `.claude/agents/specific_instructions/shared/knowledge_harvest.md` and follow
|
|
282
|
+
the protocol. Present candidates to the user for confirmation before writing.
|
|
283
|
+
|
|
284
|
+
::GATE:: id=deep-learning-engineer-phase-5 phase=5 kind=final
|
|
285
|
+
Read Phase 5 summary to the user. Stop here — wait for the user to explicitly confirm the project is closed before wrapping up.
|
|
286
|
+
::ENDGATE::
|
|
287
|
+
|
|
288
|
+
---
|
|
289
|
+
|
|
290
|
+
## When this gate is confirmed
|
|
291
|
+
|
|
292
|
+
This is the final phase. Once Syn returns APPROVED sign-off, the project is complete. Do not read further files.
|