@proflandrigan/shards 1.1.0
This diff represents the content of publicly available package versions that have been released to one of the supported registries. The information contained in this diff is provided for informational purposes only and reflects changes between package versions as they appear in their respective public registries.
- package/README.md +475 -0
- package/package.json +37 -0
- package/src/agents/academic.md +276 -0
- package/src/agents/ai-engineer.md +377 -0
- package/src/agents/analytics-engineer.md +364 -0
- package/src/agents/applied-ml-scientist.md +410 -0
- package/src/agents/backend-engineer.md +255 -0
- package/src/agents/bi-engineer.md +333 -0
- package/src/agents/data-analyst.md +343 -0
- package/src/agents/data-engineer.md +260 -0
- package/src/agents/data-modeller.md +386 -0
- package/src/agents/data-scientist.md +366 -0
- package/src/agents/deep-learning-engineer.md +389 -0
- package/src/agents/ml-engineer.md +424 -0
- package/src/agents/mlops-engineer.md +339 -0
- package/src/agents/researcher.md +187 -0
- package/src/agents/specific_instructions/academic/critical_review.md +263 -0
- package/src/agents/specific_instructions/academic/report.md +113 -0
- package/src/agents/specific_instructions/ai_engineer/advise.md +162 -0
- package/src/agents/specific_instructions/ai_engineer/bi_engineer_handoff.md +86 -0
- package/src/agents/specific_instructions/ai_engineer/experiment.md +471 -0
- package/src/agents/specific_instructions/ai_engineer/experiment_ui_mode.md +44 -0
- package/src/agents/specific_instructions/ai_engineer/phases/index.md +45 -0
- package/src/agents/specific_instructions/ai_engineer/phases/phase-1.md +55 -0
- package/src/agents/specific_instructions/ai_engineer/phases/phase-2.md +86 -0
- package/src/agents/specific_instructions/ai_engineer/phases/phase-3.md +96 -0
- package/src/agents/specific_instructions/ai_engineer/phases/phase-4.md +138 -0
- package/src/agents/specific_instructions/ai_engineer/phases/phase-5.md +157 -0
- package/src/agents/specific_instructions/ai_engineer/phases/phase-6.md +196 -0
- package/src/agents/specific_instructions/ai_engineer/phases/phase-7.md +313 -0
- package/src/agents/specific_instructions/ai_engineer/phases.md +1011 -0
- package/src/agents/specific_instructions/ai_engineer/prompt_lab.md +161 -0
- package/src/agents/specific_instructions/ai_engineer/prompt_lab_ui_mode.md +28 -0
- package/src/agents/specific_instructions/ai_engineer/research.md +393 -0
- package/src/agents/specific_instructions/ai_engineer/research_ui_mode.md +66 -0
- package/src/agents/specific_instructions/ai_engineer/review.md +159 -0
- package/src/agents/specific_instructions/ai_engineer/validation_checklist.md +182 -0
- package/src/agents/specific_instructions/analytics_engineer/advise.md +155 -0
- package/src/agents/specific_instructions/analytics_engineer/bi_engineer_handoff.md +91 -0
- package/src/agents/specific_instructions/analytics_engineer/data_analyst_handoff.md +84 -0
- package/src/agents/specific_instructions/analytics_engineer/deep_phases.md +818 -0
- package/src/agents/specific_instructions/analytics_engineer/phases_deep/index.md +24 -0
- package/src/agents/specific_instructions/analytics_engineer/phases_deep/phase-1.md +77 -0
- package/src/agents/specific_instructions/analytics_engineer/phases_deep/phase-2.md +106 -0
- package/src/agents/specific_instructions/analytics_engineer/phases_deep/phase-3.md +93 -0
- package/src/agents/specific_instructions/analytics_engineer/phases_deep/phase-4.md +79 -0
- package/src/agents/specific_instructions/analytics_engineer/phases_deep/phase-5.md +61 -0
- package/src/agents/specific_instructions/analytics_engineer/phases_deep/phase-6.md +45 -0
- package/src/agents/specific_instructions/analytics_engineer/phases_deep/phase-7.md +235 -0
- package/src/agents/specific_instructions/analytics_engineer/phases_deep/phase-8.md +221 -0
- package/src/agents/specific_instructions/analytics_engineer/phases_quick/index.md +19 -0
- package/src/agents/specific_instructions/analytics_engineer/phases_quick/phase-1.md +47 -0
- package/src/agents/specific_instructions/analytics_engineer/phases_quick/phase-2.md +78 -0
- package/src/agents/specific_instructions/analytics_engineer/quick_phases.md +112 -0
- package/src/agents/specific_instructions/analytics_engineer/review.md +167 -0
- package/src/agents/specific_instructions/analytics_engineer/service_mode.md +369 -0
- package/src/agents/specific_instructions/analytics_engineer/ui_mode.md +45 -0
- package/src/agents/specific_instructions/analytics_engineer/update.md +162 -0
- package/src/agents/specific_instructions/analytics_engineer/validation_checklist.md +121 -0
- package/src/agents/specific_instructions/applied_ml_scientist/advise.md +143 -0
- package/src/agents/specific_instructions/applied_ml_scientist/phases/index.md +21 -0
- package/src/agents/specific_instructions/applied_ml_scientist/phases/phase-1.md +51 -0
- package/src/agents/specific_instructions/applied_ml_scientist/phases/phase-2.md +66 -0
- package/src/agents/specific_instructions/applied_ml_scientist/phases/phase-3.md +113 -0
- package/src/agents/specific_instructions/applied_ml_scientist/phases/phase-4.md +104 -0
- package/src/agents/specific_instructions/applied_ml_scientist/phases/phase-5.md +156 -0
- package/src/agents/specific_instructions/applied_ml_scientist/phases.md +428 -0
- package/src/agents/specific_instructions/applied_ml_scientist/research.md +379 -0
- package/src/agents/specific_instructions/applied_ml_scientist/review.md +142 -0
- package/src/agents/specific_instructions/applied_ml_scientist/validation_checklist.md +136 -0
- package/src/agents/specific_instructions/backend_engineer/clean.md +149 -0
- package/src/agents/specific_instructions/backend_engineer/review.md +91 -0
- package/src/agents/specific_instructions/backend_engineer/review_checklist.md +54 -0
- package/src/agents/specific_instructions/backend_engineer/service_mode.md +67 -0
- package/src/agents/specific_instructions/bi_engineer/advise.md +137 -0
- package/src/agents/specific_instructions/bi_engineer/data_analyst_handoff.md +77 -0
- package/src/agents/specific_instructions/bi_engineer/incoming_handoff.md +45 -0
- package/src/agents/specific_instructions/bi_engineer/phases/index.md +20 -0
- package/src/agents/specific_instructions/bi_engineer/phases/phase-1.md +164 -0
- package/src/agents/specific_instructions/bi_engineer/phases/phase-2.md +92 -0
- package/src/agents/specific_instructions/bi_engineer/phases/phase-3.md +121 -0
- package/src/agents/specific_instructions/bi_engineer/phases/phase-4.md +106 -0
- package/src/agents/specific_instructions/bi_engineer/phases.md +451 -0
- package/src/agents/specific_instructions/bi_engineer/review.md +166 -0
- package/src/agents/specific_instructions/bi_engineer/update.md +147 -0
- package/src/agents/specific_instructions/bi_engineer/validation_checklist.md +124 -0
- package/src/agents/specific_instructions/data_analyst/advise.md +138 -0
- package/src/agents/specific_instructions/data_analyst/explain.md +221 -0
- package/src/agents/specific_instructions/data_analyst/incoming_handoff.md +40 -0
- package/src/agents/specific_instructions/data_analyst/phases/index.md +20 -0
- package/src/agents/specific_instructions/data_analyst/phases/phase-1.md +159 -0
- package/src/agents/specific_instructions/data_analyst/phases/phase-2.md +112 -0
- package/src/agents/specific_instructions/data_analyst/phases/phase-3.md +265 -0
- package/src/agents/specific_instructions/data_analyst/phases/phase-4.md +100 -0
- package/src/agents/specific_instructions/data_analyst/phases.md +501 -0
- package/src/agents/specific_instructions/data_analyst/review.md +138 -0
- package/src/agents/specific_instructions/data_analyst/ui_mode.md +26 -0
- package/src/agents/specific_instructions/data_analyst/update.md +144 -0
- package/src/agents/specific_instructions/data_analyst/validation_checklist.md +95 -0
- package/src/agents/specific_instructions/data_engineer/advise.md +137 -0
- package/src/agents/specific_instructions/data_engineer/phases.md +466 -0
- package/src/agents/specific_instructions/data_engineer/phases_deep/index.md +23 -0
- package/src/agents/specific_instructions/data_engineer/phases_deep/phase-1.md +49 -0
- package/src/agents/specific_instructions/data_engineer/phases_deep/phase-2.md +93 -0
- package/src/agents/specific_instructions/data_engineer/phases_deep/phase-3.md +55 -0
- package/src/agents/specific_instructions/data_engineer/phases_deep/phase-4.md +48 -0
- package/src/agents/specific_instructions/data_engineer/phases_deep/phase-5.md +40 -0
- package/src/agents/specific_instructions/data_engineer/phases_deep/phase-6.md +102 -0
- package/src/agents/specific_instructions/data_engineer/phases_deep/phase-7.md +87 -0
- package/src/agents/specific_instructions/data_engineer/phases_quick/index.md +19 -0
- package/src/agents/specific_instructions/data_engineer/phases_quick/phase-1.md +45 -0
- package/src/agents/specific_instructions/data_engineer/phases_quick/phase-2.md +54 -0
- package/src/agents/specific_instructions/data_engineer/review.md +135 -0
- package/src/agents/specific_instructions/data_engineer/validation_checklist.md +136 -0
- package/src/agents/specific_instructions/data_modeller/advise.md +137 -0
- package/src/agents/specific_instructions/data_modeller/phases.md +581 -0
- package/src/agents/specific_instructions/data_modeller/phases_deep/index.md +23 -0
- package/src/agents/specific_instructions/data_modeller/phases_deep/phase-1.md +52 -0
- package/src/agents/specific_instructions/data_modeller/phases_deep/phase-2.md +113 -0
- package/src/agents/specific_instructions/data_modeller/phases_deep/phase-3.md +47 -0
- package/src/agents/specific_instructions/data_modeller/phases_deep/phase-4.md +51 -0
- package/src/agents/specific_instructions/data_modeller/phases_deep/phase-5.md +45 -0
- package/src/agents/specific_instructions/data_modeller/phases_deep/phase-6.md +105 -0
- package/src/agents/specific_instructions/data_modeller/phases_deep/phase-7.md +136 -0
- package/src/agents/specific_instructions/data_modeller/phases_quick/index.md +19 -0
- package/src/agents/specific_instructions/data_modeller/phases_quick/phase-1.md +47 -0
- package/src/agents/specific_instructions/data_modeller/phases_quick/phase-2.md +65 -0
- package/src/agents/specific_instructions/data_modeller/review.md +141 -0
- package/src/agents/specific_instructions/data_modeller/service_mode.md +218 -0
- package/src/agents/specific_instructions/data_modeller/validation_checklist.md +125 -0
- package/src/agents/specific_instructions/data_scientist/advise.md +158 -0
- package/src/agents/specific_instructions/data_scientist/bi_engineer_handoff.md +63 -0
- package/src/agents/specific_instructions/data_scientist/experiment.md +482 -0
- package/src/agents/specific_instructions/data_scientist/experiment_ui_mode.md +44 -0
- package/src/agents/specific_instructions/data_scientist/explain.md +247 -0
- package/src/agents/specific_instructions/data_scientist/greenfield_data.md +35 -0
- package/src/agents/specific_instructions/data_scientist/ml_engineer_handoff.md +52 -0
- package/src/agents/specific_instructions/data_scientist/notebook_walkthrough.md +76 -0
- package/src/agents/specific_instructions/data_scientist/phases/index.md +24 -0
- package/src/agents/specific_instructions/data_scientist/phases/phase-1.md +45 -0
- package/src/agents/specific_instructions/data_scientist/phases/phase-2.md +67 -0
- package/src/agents/specific_instructions/data_scientist/phases/phase-3.md +89 -0
- package/src/agents/specific_instructions/data_scientist/phases/phase-4.md +143 -0
- package/src/agents/specific_instructions/data_scientist/phases/phase-5.md +71 -0
- package/src/agents/specific_instructions/data_scientist/phases/phase-6.md +239 -0
- package/src/agents/specific_instructions/data_scientist/phases/phase-7.md +207 -0
- package/src/agents/specific_instructions/data_scientist/phases.md +651 -0
- package/src/agents/specific_instructions/data_scientist/research.md +345 -0
- package/src/agents/specific_instructions/data_scientist/research_ui_mode.md +52 -0
- package/src/agents/specific_instructions/data_scientist/review.md +136 -0
- package/src/agents/specific_instructions/data_scientist/service_mode.md +247 -0
- package/src/agents/specific_instructions/data_scientist/validation_checklist.md +183 -0
- package/src/agents/specific_instructions/deep_learning_engineer/advise.md +145 -0
- package/src/agents/specific_instructions/deep_learning_engineer/phases/index.md +21 -0
- package/src/agents/specific_instructions/deep_learning_engineer/phases/phase-1.md +74 -0
- package/src/agents/specific_instructions/deep_learning_engineer/phases/phase-2.md +98 -0
- package/src/agents/specific_instructions/deep_learning_engineer/phases/phase-3.md +76 -0
- package/src/agents/specific_instructions/deep_learning_engineer/phases/phase-4.md +128 -0
- package/src/agents/specific_instructions/deep_learning_engineer/phases/phase-5.md +292 -0
- package/src/agents/specific_instructions/deep_learning_engineer/phases.md +567 -0
- package/src/agents/specific_instructions/deep_learning_engineer/research.md +389 -0
- package/src/agents/specific_instructions/deep_learning_engineer/review.md +155 -0
- package/src/agents/specific_instructions/deep_learning_engineer/validation_checklist.md +147 -0
- package/src/agents/specific_instructions/ml_engineer/advise.md +174 -0
- package/src/agents/specific_instructions/ml_engineer/bi_engineer_handoff.md +71 -0
- package/src/agents/specific_instructions/ml_engineer/experiment.md +474 -0
- package/src/agents/specific_instructions/ml_engineer/experiment_ui_mode.md +44 -0
- package/src/agents/specific_instructions/ml_engineer/notebook_walkthrough.md +75 -0
- package/src/agents/specific_instructions/ml_engineer/phases/index.md +25 -0
- package/src/agents/specific_instructions/ml_engineer/phases/phase-1.md +49 -0
- package/src/agents/specific_instructions/ml_engineer/phases/phase-2.md +75 -0
- package/src/agents/specific_instructions/ml_engineer/phases/phase-3.md +124 -0
- package/src/agents/specific_instructions/ml_engineer/phases/phase-4.md +279 -0
- package/src/agents/specific_instructions/ml_engineer/phases/phase-5.md +160 -0
- package/src/agents/specific_instructions/ml_engineer/phases/phase-6-5.md +170 -0
- package/src/agents/specific_instructions/ml_engineer/phases/phase-6.md +295 -0
- package/src/agents/specific_instructions/ml_engineer/phases/phase-7.md +337 -0
- package/src/agents/specific_instructions/ml_engineer/phases.md +1068 -0
- package/src/agents/specific_instructions/ml_engineer/research.md +437 -0
- package/src/agents/specific_instructions/ml_engineer/research_ui_mode.md +71 -0
- package/src/agents/specific_instructions/ml_engineer/review.md +187 -0
- package/src/agents/specific_instructions/ml_engineer/service_mode.md +273 -0
- package/src/agents/specific_instructions/ml_engineer/validation_checklist.md +185 -0
- package/src/agents/specific_instructions/mlops_engineer/advise.md +139 -0
- package/src/agents/specific_instructions/mlops_engineer/phases/index.md +23 -0
- package/src/agents/specific_instructions/mlops_engineer/phases/phase-1.md +52 -0
- package/src/agents/specific_instructions/mlops_engineer/phases/phase-2.md +86 -0
- package/src/agents/specific_instructions/mlops_engineer/phases/phase-3.md +105 -0
- package/src/agents/specific_instructions/mlops_engineer/phases/phase-4.md +128 -0
- package/src/agents/specific_instructions/mlops_engineer/phases/phase-5.md +106 -0
- package/src/agents/specific_instructions/mlops_engineer/phases/phase-6.md +128 -0
- package/src/agents/specific_instructions/mlops_engineer/phases/phase-7.md +144 -0
- package/src/agents/specific_instructions/mlops_engineer/phases.md +671 -0
- package/src/agents/specific_instructions/mlops_engineer/review.md +164 -0
- package/src/agents/specific_instructions/mlops_engineer/service_mode.md +81 -0
- package/src/agents/specific_instructions/mlops_engineer/validation_checklist.md +151 -0
- package/src/agents/specific_instructions/researcher/critical_review.md +292 -0
- package/src/agents/specific_instructions/researcher/review_checklist.md +67 -0
- package/src/agents/specific_instructions/researcher/service_mode.md +224 -0
- package/src/agents/specific_instructions/shared/auto_verify_mode.md +141 -0
- package/src/agents/specific_instructions/shared/autonomous_research.md +1289 -0
- package/src/agents/specific_instructions/shared/behavioral_rules.md +36 -0
- package/src/agents/specific_instructions/shared/diverge_protocol.md +387 -0
- package/src/agents/specific_instructions/shared/engineering_guidelines.md +136 -0
- package/src/agents/specific_instructions/shared/experiment_versioning.md +184 -0
- package/src/agents/specific_instructions/shared/goal_mode.md +187 -0
- package/src/agents/specific_instructions/shared/incremental_testing.md +139 -0
- package/src/agents/specific_instructions/shared/intent_discovery.md +223 -0
- package/src/agents/specific_instructions/shared/join_path_protocol.md +168 -0
- package/src/agents/specific_instructions/shared/knowledge_checkpoint.md +83 -0
- package/src/agents/specific_instructions/shared/knowledge_harvest.md +220 -0
- package/src/agents/specific_instructions/shared/knowledge_retrieval.md +100 -0
- package/src/agents/specific_instructions/shared/notebook_walkthrough_protocol.md +367 -0
- package/src/agents/specific_instructions/shared/reviewer_verdict_protocol.md +74 -0
- package/src/agents/specific_instructions/shared/swarm_protocol.md +97 -0
- package/src/agents/specific_instructions/shared/validation_protocol.md +139 -0
- package/src/agents/specific_instructions/syn/arbiter.md +140 -0
- package/src/agents/specific_instructions/syn/brainstorm.md +550 -0
- package/src/agents/specific_instructions/syn/code_review.md +232 -0
- package/src/agents/specific_instructions/syn/diff.md +239 -0
- package/src/agents/specific_instructions/syn/final_review.md +65 -0
- package/src/agents/specific_instructions/syn/fixer.md +240 -0
- package/src/agents/specific_instructions/syn/free_form.md +130 -0
- package/src/agents/specific_instructions/syn/knowledge.md +468 -0
- package/src/agents/specific_instructions/syn/notebook_walkthrough.md +78 -0
- package/src/agents/specific_instructions/syn/panel_review.md +634 -0
- package/src/agents/specific_instructions/syn/pm.md +453 -0
- package/src/agents/specific_instructions/syn/pr_review.md +255 -0
- package/src/agents/specific_instructions/syn/slides.md +417 -0
- package/src/agents/syn.md +729 -0
- package/src/commands/academic.md +41 -0
- package/src/commands/ai-engineer.md +45 -0
- package/src/commands/analytics-engineer.md +48 -0
- package/src/commands/applied-ml-scientist.md +45 -0
- package/src/commands/backend-engineer.md +35 -0
- package/src/commands/bi-engineer.md +40 -0
- package/src/commands/brainstorm.md +24 -0
- package/src/commands/data-analyst.md +38 -0
- package/src/commands/data-engineer.md +37 -0
- package/src/commands/data-modeller.md +38 -0
- package/src/commands/data-scientist.md +38 -0
- package/src/commands/deep-learning-engineer.md +47 -0
- package/src/commands/end.md +49 -0
- package/src/commands/knowledge.md +24 -0
- package/src/commands/ml-engineer.md +42 -0
- package/src/commands/mlops-engineer.md +47 -0
- package/src/commands/notebook-walkthrough.md +58 -0
- package/src/commands/researcher.md +40 -0
- package/src/commands/resume.md +57 -0
- package/src/commands/review-pr.md +26 -0
- package/src/commands/shards-guide.md +41 -0
- package/src/commands/shards-ui.md +32 -0
- package/src/commands/shards.md +41 -0
- package/src/docs/01-getting-started/concepts.md +109 -0
- package/src/docs/01-getting-started/first-session.md +79 -0
- package/src/docs/01-getting-started/install.md +61 -0
- package/src/docs/02-agents/academic.md +71 -0
- package/src/docs/02-agents/ai-engineer.md +78 -0
- package/src/docs/02-agents/analytics-engineer.md +58 -0
- package/src/docs/02-agents/applied-ml-scientist.md +59 -0
- package/src/docs/02-agents/backend-engineer.md +58 -0
- package/src/docs/02-agents/bi-engineer.md +65 -0
- package/src/docs/02-agents/data-analyst.md +67 -0
- package/src/docs/02-agents/data-engineer.md +57 -0
- package/src/docs/02-agents/data-modeller.md +51 -0
- package/src/docs/02-agents/data-scientist.md +78 -0
- package/src/docs/02-agents/deep-learning-engineer.md +64 -0
- package/src/docs/02-agents/ml-engineer.md +80 -0
- package/src/docs/02-agents/mlops-engineer.md +59 -0
- package/src/docs/02-agents/overview.md +62 -0
- package/src/docs/02-agents/researcher.md +73 -0
- package/src/docs/02-agents/syn.md +88 -0
- package/src/docs/03-protocols/auto-verify.md +82 -0
- package/src/docs/03-protocols/autonomous-research.md +59 -0
- package/src/docs/03-protocols/behavioral-rules.md +35 -0
- package/src/docs/03-protocols/diverge.md +50 -0
- package/src/docs/03-protocols/engineering-guidelines.md +56 -0
- package/src/docs/03-protocols/experiment-versioning.md +38 -0
- package/src/docs/03-protocols/gate-pattern.md +65 -0
- package/src/docs/03-protocols/incremental-testing.md +68 -0
- package/src/docs/03-protocols/join-path.md +46 -0
- package/src/docs/03-protocols/knowledge-ledger.md +70 -0
- package/src/docs/03-protocols/reviewer-verdicts.md +39 -0
- package/src/docs/03-protocols/swarm.md +40 -0
- package/src/docs/03-protocols/validation.md +174 -0
- package/src/docs/04-ui/activity-bar.md +70 -0
- package/src/docs/04-ui/chat-pane.md +80 -0
- package/src/docs/04-ui/code-intel.md +62 -0
- package/src/docs/04-ui/file-editing.md +61 -0
- package/src/docs/04-ui/git.md +54 -0
- package/src/docs/04-ui/keybindings.md +79 -0
- package/src/docs/04-ui/knowledge-map.md +76 -0
- package/src/docs/04-ui/overview.md +93 -0
- package/src/docs/04-ui/panels.md +49 -0
- package/src/docs/04-ui/pinboard-selection.md +66 -0
- package/src/docs/04-ui/quick-open-palette.md +56 -0
- package/src/docs/04-ui/sessions.md +81 -0
- package/src/docs/04-ui/settings-permissions.md +56 -0
- package/src/docs/05-commands/reference.md +59 -0
- package/src/docs/06-outputs/directory-map.md +116 -0
- package/src/docs/07-workflows/ai-eval-first.md +57 -0
- package/src/docs/07-workflows/deep-study-to-production.md +76 -0
- package/src/docs/07-workflows/diverge-exploration.md +77 -0
- package/src/docs/07-workflows/quick-analysis.md +45 -0
- package/src/docs/08-integrations/claude-code-auto-mode.md +191 -0
- package/src/docs/08-integrations/google-slides.md +175 -0
- package/src/docs/README.md +30 -0
- package/src/docs/manifest.json +108 -0
- package/src/templates/analysis-template.md +20 -0
- package/src/templates/branch-report.md +46 -0
- package/src/templates/diff-report.md +88 -0
- package/src/templates/knowledge-index.md +7 -0
- package/src/templates/model-card-schema.json +186 -0
- package/src/templates/model-card-schema.md +88 -0
- package/src/templates/model-card.md +124 -0
- package/src/templates/project-plan.md +47 -0
- package/src/templates/project-specs.md +81 -0
- package/src/templates/report-template.md +43 -0
- package/src/templates/study-template.md +25 -0
- package/src/ui/cc-readonly.js +181 -0
- package/src/ui/chat-session.js +466 -0
- package/src/ui/css/base.css +136 -0
- package/src/ui/css/brainstorm.css +525 -0
- package/src/ui/css/chat.css +1405 -0
- package/src/ui/css/editor.css +546 -0
- package/src/ui/css/eval-dashboard.css +157 -0
- package/src/ui/css/experiment.css +237 -0
- package/src/ui/css/guide.css +186 -0
- package/src/ui/css/knowledge-map.css +383 -0
- package/src/ui/css/layout.css +431 -0
- package/src/ui/css/model-card.css +161 -0
- package/src/ui/css/notebook-walkthrough.css +271 -0
- package/src/ui/css/pr-review.css +403 -0
- package/src/ui/css/prompt-lab.css +325 -0
- package/src/ui/css/sessions.css +258 -0
- package/src/ui/css/sidebar.css +661 -0
- package/src/ui/css/terminal.css +113 -0
- package/src/ui/css/theme-light.css +542 -0
- package/src/ui/index.html +389 -0
- package/src/ui/js/agents.js +32 -0
- package/src/ui/js/bookmarks.js +230 -0
- package/src/ui/js/chat.js +1776 -0
- package/src/ui/js/code-intel.js +328 -0
- package/src/ui/js/command-palette.js +142 -0
- package/src/ui/js/events.js +591 -0
- package/src/ui/js/explorer.js +317 -0
- package/src/ui/js/file-view.js +477 -0
- package/src/ui/js/git.js +536 -0
- package/src/ui/js/guide.js +198 -0
- package/src/ui/js/hud.js +75 -0
- package/src/ui/js/init.js +351 -0
- package/src/ui/js/knowledge-map.js +906 -0
- package/src/ui/js/markdown.js +114 -0
- package/src/ui/js/monaco.js +164 -0
- package/src/ui/js/notebook-walkthrough.js +272 -0
- package/src/ui/js/notebook.js +448 -0
- package/src/ui/js/panels.js +2681 -0
- package/src/ui/js/pinboard.js +186 -0
- package/src/ui/js/quick-open.js +164 -0
- package/src/ui/js/selection-context.js +131 -0
- package/src/ui/js/sessions.js +256 -0
- package/src/ui/js/settings.js +476 -0
- package/src/ui/js/split-view.js +82 -0
- package/src/ui/js/state.js +343 -0
- package/src/ui/js/table.js +161 -0
- package/src/ui/js/tabs.js +284 -0
- package/src/ui/js/tabular.js +125 -0
- package/src/ui/js/terminal.js +354 -0
- package/src/ui/js/timeline.js +137 -0
- package/src/ui/js/utils.js +293 -0
- package/src/ui/notebook-kernel.py +790 -0
- package/src/ui/open-browser.js +55 -0
- package/src/ui/permission-pattern.js +42 -0
- package/src/ui/relay.js +513 -0
- package/src/ui/server.js +3072 -0
- package/src/ui/session-index.js +225 -0
- package/src/ui/shards_icon.png +0 -0
- package/src/ui/spawn-server.js +41 -0
- package/src/ui/symbol-index.js +813 -0
- package/src/ui/ui-push.js +177 -0
- package/tools/gate-hook/VALIDATION_SPEC.md +273 -0
- package/tools/gate-hook/__tests__/auto-verify.test.js +343 -0
- package/tools/gate-hook/auto-allowlist.js +179 -0
- package/tools/gate-hook/auto-state.js +68 -0
- package/tools/gate-hook/classify.js +21 -0
- package/tools/gate-hook/log.js +57 -0
- package/tools/gate-hook/parser.js +205 -0
- package/tools/gate-hook/sql-guard.js +230 -0
- package/tools/gate-hook/state.js +170 -0
- package/tools/gate-hook/sweep.js +139 -0
- package/tools/gate-hook/transcript.js +45 -0
- package/tools/gate-hook/validation.js +321 -0
- package/tools/gate-hook.js +475 -0
- package/tools/install.js +914 -0
- package/tools/shards-gates.js +311 -0
- package/tools/shards-sessions.js +261 -0
- package/tools/shards-ui.js +377 -0
|
@@ -0,0 +1,345 @@
|
|
|
1
|
+
# Data Scientist Autonomous Research Mode
|
|
2
|
+
|
|
3
|
+
This file governs `[AR]` — Autonomous Research mode for the Data Scientist. A
|
|
4
|
+
self-steering loop that iteratively pushes a single primary metric on a
|
|
5
|
+
modeling or analytical problem, generating hypotheses adaptively, auto-keeping
|
|
6
|
+
or auto-reverting each change.
|
|
7
|
+
|
|
8
|
+
You are the Data Scientist throughout. No persona transfer. You remain
|
|
9
|
+
condescending and particular. Since you're now iterating autonomously, be
|
|
10
|
+
especially particular about causal honesty — correlation is still not
|
|
11
|
+
causation even when the AR loop is running.
|
|
12
|
+
|
|
13
|
+
Read `.claude/agents/specific_instructions/shared/autonomous_research.md` in
|
|
14
|
+
full before executing this file.
|
|
15
|
+
|
|
16
|
+
---
|
|
17
|
+
|
|
18
|
+
## When to use `[AR]` vs `[EXP]`
|
|
19
|
+
|
|
20
|
+
| Mode | Shape | Use when |
|
|
21
|
+
|------|-------|----------|
|
|
22
|
+
| `[EXP]` | 3-5 pre-planned experiments on an existing study | You know what to try |
|
|
23
|
+
| `[AR]` interactive | 10 adaptive iterations against a metric | You want to push a modeling or feature-engineering result, conversationally |
|
|
24
|
+
| `[AR]` overnight | 100 adaptive iterations | You want a long AR session against a single metric |
|
|
25
|
+
| `[AR]` fan-out | K parallel AR loops per approach family | You want methodology alternatives compared head-to-head |
|
|
26
|
+
|
|
27
|
+
---
|
|
28
|
+
|
|
29
|
+
## Phase 0 — Research Setup (GATE)
|
|
30
|
+
|
|
31
|
+
### Context loading
|
|
32
|
+
|
|
33
|
+
1. Locate `project-specs.md` at `studies/<project_name>/project-specs.md` or
|
|
34
|
+
the existing study directory.
|
|
35
|
+
2. Read `project-specs.md` in full.
|
|
36
|
+
3. Scan for relevant artifacts: queries, notebooks, feature engineering code,
|
|
37
|
+
model configs, evaluation output.
|
|
38
|
+
4. Identify the current metrics baseline.
|
|
39
|
+
5. Establish `<study_dir>/experiments/`.
|
|
40
|
+
|
|
41
|
+
### Versioning detection
|
|
42
|
+
|
|
43
|
+
Per `experiment_versioning.md` Section A. AR requires git (or DVC). If
|
|
44
|
+
versioning is `none`, warn and offer to init, drop to `[EXP]`, or cancel.
|
|
45
|
+
|
|
46
|
+
### Knowledge retrieval
|
|
47
|
+
|
|
48
|
+
Per `knowledge_retrieval.md` AR entry point. Match on metric, domain, and
|
|
49
|
+
feature/methodology family.
|
|
50
|
+
|
|
51
|
+
### Preset selection
|
|
52
|
+
|
|
53
|
+
```
|
|
54
|
+
AR runs in one of two presets:
|
|
55
|
+
|
|
56
|
+
[interactive] — budget=10, reviewer cadence=3. You're nearby.
|
|
57
|
+
[overnight] — budget=100, reviewer cadence=10, cost ceiling required.
|
|
58
|
+
Interrupt anytime by editing experiments/research_brief.md
|
|
59
|
+
Steering Notes (re-read every iteration).
|
|
60
|
+
[custom] — I ask you for each parameter.
|
|
61
|
+
```
|
|
62
|
+
|
|
63
|
+
### Parameter confirmation
|
|
64
|
+
|
|
65
|
+
- **Primary metric:** single north-star. Examples: AUC, log-loss, RMSE, MAE,
|
|
66
|
+
F1, R², lift@k, effect-size magnitude (for causal studies).
|
|
67
|
+
- **Direction:** maximize | minimize
|
|
68
|
+
- **Baseline + source**
|
|
69
|
+
- **Target** (optional)
|
|
70
|
+
- **Iteration budget**
|
|
71
|
+
- **Per-iteration time limit**
|
|
72
|
+
- **Max consecutive regressions** (default: 3)
|
|
73
|
+
- **Metric degradation floor** (optional — especially important on causal
|
|
74
|
+
studies where a regression may indicate the model fit to noise)
|
|
75
|
+
- **Epsilon** (default: 1% of baseline)
|
|
76
|
+
- **Cost ceiling:** required for overnight
|
|
77
|
+
- **Reviewer cadence** (default: 3 interactive / 10 overnight)
|
|
78
|
+
- **Plateau window W** (default: 5)
|
|
79
|
+
- **Diminishing returns threshold** (default: 0.1% of baseline)
|
|
80
|
+
- **Full eval cadence M** (default: 5 interactive / 10 overnight)
|
|
81
|
+
- **Mutable scope:**
|
|
82
|
+
- Typical: `notebooks/`, `queries/`, feature engineering code, model config
|
|
83
|
+
- For causal studies: mutable is typically model spec, not query logic
|
|
84
|
+
- **Immutable scope:**
|
|
85
|
+
- Raw data, eval harness, study spec itself (`project-specs.md`)
|
|
86
|
+
- For causal studies: any variable included as an outcome or treatment
|
|
87
|
+
|
|
88
|
+
### UI detection
|
|
89
|
+
|
|
90
|
+
If `.shards/ui.port` exists, read
|
|
91
|
+
`.claude/agents/specific_instructions/data_scientist/research_ui_mode.md` in full.
|
|
92
|
+
|
|
93
|
+
### Document Phase 0
|
|
94
|
+
|
|
95
|
+
Append to `project-specs.md`:
|
|
96
|
+
|
|
97
|
+
```markdown
|
|
98
|
+
---
|
|
99
|
+
|
|
100
|
+
## Phase 0: AR Setup (Data Scientist)
|
|
101
|
+
|
|
102
|
+
- **Mode:** Autonomous Research (`[AR]`)
|
|
103
|
+
- **Preset:** <interactive | overnight | custom>
|
|
104
|
+
- **Study type:** <predictive modeling | causal study | EDA-driven exploration | other>
|
|
105
|
+
- **Primary metric:** <name> (<direction>)
|
|
106
|
+
- **Baseline:** <value> (source: <source>)
|
|
107
|
+
- **Target:** <value or "none">
|
|
108
|
+
- **Iteration budget:** <N>
|
|
109
|
+
- **Reviewer cadence:** <K>
|
|
110
|
+
- **Cost ceiling:** <tokens: N / dollars: N, or "none">
|
|
111
|
+
- **Metric floor:** <value or "none">
|
|
112
|
+
- **Mutable scope:** <list>
|
|
113
|
+
- **Immutable scope:** <list>
|
|
114
|
+
- **Versioning mode:** <git | dvc>
|
|
115
|
+
|
|
116
|
+
### Knowledge Ledger
|
|
117
|
+
- **Entries checked:** <N>
|
|
118
|
+
- **Relevant entries found:** <N>
|
|
119
|
+
- <title> (<type>, <confidence>) — <relevance>
|
|
120
|
+
- **Or:** No relevant entries found
|
|
121
|
+
- **Relevant features:** <N>
|
|
122
|
+
```
|
|
123
|
+
|
|
124
|
+
::GATE:: id=specific-instructions-data-scientist-research-phase0 phase=0 kind=execute
|
|
125
|
+
Read this section back. Stop here. Wait for confirmation.
|
|
126
|
+
::ENDGATE::
|
|
127
|
+
|
|
128
|
+
---
|
|
129
|
+
|
|
130
|
+
## Phase 1 — Research Brief + Optional DIVERGE (GATE)
|
|
131
|
+
|
|
132
|
+
### Draft the research brief
|
|
133
|
+
|
|
134
|
+
Follow Section A of `autonomous_research.md`. Use
|
|
135
|
+
`templates/research-brief.md`, write to `<study_dir>/experiments/research_brief.md`.
|
|
136
|
+
Write `results.json` with `mode: "autonomous-research"`.
|
|
137
|
+
|
|
138
|
+
Update `project-specs.md` with a `## Autonomous Research` section.
|
|
139
|
+
|
|
140
|
+
### Consider DIVERGE fan-out
|
|
141
|
+
|
|
142
|
+
**Typical Data Scientist approach families for fan-out:**
|
|
143
|
+
- Tree-based (gradient boosting family)
|
|
144
|
+
- Linear (regularized regression, logistic, GLM)
|
|
145
|
+
- Neural (simple MLP or deeper — if data supports)
|
|
146
|
+
- Causal methods (DiD, instrumental variables, synthetic control) — only if
|
|
147
|
+
the study is causal
|
|
148
|
+
- Feature-heavy vs. model-heavy (same model family, different feature
|
|
149
|
+
strategies)
|
|
150
|
+
|
|
151
|
+
**Typical slugs:** `ds-gbdt`, `ds-linear`, `ds-neural`, `ds-features-heavy`.
|
|
152
|
+
|
|
153
|
+
Propose DIVERGE per `diverge_protocol.md` Section B with AR gate ID namespace.
|
|
154
|
+
|
|
155
|
+
### Behavioral exception announcement
|
|
156
|
+
|
|
157
|
+
Before the gate:
|
|
158
|
+
|
|
159
|
+
> "Facilitate, don't generate" is suspended for Phase 2. I'll autonomously
|
|
160
|
+
> generate hypotheses, implement changes, and auto-decide keep/revert. This
|
|
161
|
+
> means I'm running feature engineering and model training without checking
|
|
162
|
+
> in with you between iterations. You can steer at any time by editing
|
|
163
|
+
> `experiments/research_brief.md` — I re-read it every iteration. Phase 0,
|
|
164
|
+
> Phase 1, and Phase 3 remain gated.
|
|
165
|
+
|
|
166
|
+
### Optional `/goal` activation
|
|
167
|
+
|
|
168
|
+
Read `.claude/agents/specific_instructions/shared/goal_mode.md` in full before
|
|
169
|
+
writing the gate. Compose a candidate `/goal` condition from this run's
|
|
170
|
+
Phase 0 settings (primary metric, direction, target if set, iteration budget,
|
|
171
|
+
metric floor) using the AR condition template, and include the resulting
|
|
172
|
+
copy-paste block in the message that precedes the Phase 1 gate:
|
|
173
|
+
|
|
174
|
+
```text
|
|
175
|
+
/goal The AR loop is complete when ANY of the following is true:
|
|
176
|
+
(a) the most recent inline iteration summary shows <primary_metric> has
|
|
177
|
+
<crossed target X in the maximize direction
|
|
178
|
+
| dropped below target X in the minimize direction>;
|
|
179
|
+
(b) the most recent iteration summary or status line contains
|
|
180
|
+
"Convergence detected" with reason in {plateau, diminishing-returns,
|
|
181
|
+
budget-exhausted, cost-ceiling, consecutive-failures,
|
|
182
|
+
metric-floor-breach, user-interrupt, reviewer-pause,
|
|
183
|
+
scope-violation, error-limit, timeout-limit};
|
|
184
|
+
(c) the agent has begun writing the Phase 3 research summary
|
|
185
|
+
(look for "Phase 3" or "research_summary.md").
|
|
186
|
+
Or stop after <budget+5> turns.
|
|
187
|
+
```
|
|
188
|
+
|
|
189
|
+
If no target was set, drop clause (a). Activation is optional:
|
|
190
|
+
- **With `/goal`:** Phase 2 runs without per-iteration prompts. Transcript
|
|
191
|
+
discipline (`autonomous_research.md` §B.4/B.8) is mandatory — the evaluator
|
|
192
|
+
reads only the conversation, not files.
|
|
193
|
+
- **Without `/goal`:** §E convergence and §G safety rails still terminate
|
|
194
|
+
the loop. Per-iteration echoes remain recommended for readability.
|
|
195
|
+
|
|
196
|
+
If `/goal` is unavailable (Code < v2.1.139, `disableAllHooks` set, command
|
|
197
|
+
rejected), accept that and proceed — the loop still runs and terminates per
|
|
198
|
+
the existing logic.
|
|
199
|
+
|
|
200
|
+
### Gate
|
|
201
|
+
|
|
202
|
+
::GATE:: id=specific-instructions-data-scientist-research-phase1 phase=1 kind=execute
|
|
203
|
+
Read the brief back. Last checkpoint before autonomous execution. Wait for
|
|
204
|
+
explicit confirmation.
|
|
205
|
+
::ENDGATE::
|
|
206
|
+
|
|
207
|
+
---
|
|
208
|
+
|
|
209
|
+
## Phase 2 — Autonomous Research Loop (NO GATES by default)
|
|
210
|
+
|
|
211
|
+
Follow Section B of `autonomous_research.md`.
|
|
212
|
+
|
|
213
|
+
### Reviewer: Researcher
|
|
214
|
+
|
|
215
|
+
Your primary reviewer for AR is the **Researcher** — statistical methodology,
|
|
216
|
+
assumption validation, outlier handling, distribution claims. The Researcher
|
|
217
|
+
is consulted via Task per Section D.4.
|
|
218
|
+
|
|
219
|
+
Secondary reviewer ad hoc: **ML Engineer** when an iteration touches
|
|
220
|
+
production-relevant modeling (latency, model size, serving-feasibility).
|
|
221
|
+
|
|
222
|
+
Standard cadence:
|
|
223
|
+
- Always first iteration
|
|
224
|
+
- Every K iterations
|
|
225
|
+
- After improvements > 5% of baseline
|
|
226
|
+
- Before stopping on consecutive regression limit
|
|
227
|
+
- When Steering Notes change
|
|
228
|
+
|
|
229
|
+
AR-specific verdicts: `CONTINUE`, `REDIRECT`, `PAUSE`, `RETRO_REVERT`.
|
|
230
|
+
|
|
231
|
+
The Researcher is especially likely to return `RETRO_REVERT` for:
|
|
232
|
+
- Data leakage (a feature computed using post-outcome information)
|
|
233
|
+
- Target leakage (a feature that is a near-deterministic function of the target)
|
|
234
|
+
- Temporal leakage (cross-validation folds that ignore time ordering)
|
|
235
|
+
- Sample contamination (training on examples that appear in holdout)
|
|
236
|
+
|
|
237
|
+
Apply `reviewer_verdict_protocol.md` afterward.
|
|
238
|
+
|
|
239
|
+
### Hypothesis categories for Data Scientist
|
|
240
|
+
|
|
241
|
+
Draw from these adaptively:
|
|
242
|
+
|
|
243
|
+
**Feature engineering**
|
|
244
|
+
- New features: interactions, lags, rolling aggregations, ratios
|
|
245
|
+
- Transformations: log, sqrt, Box-Cox, binning
|
|
246
|
+
- Remove leaky / target-contaminated features
|
|
247
|
+
- Encoding: target encoding, mean encoding, embedding-based
|
|
248
|
+
|
|
249
|
+
**Model family**
|
|
250
|
+
- Linear / regularized (logistic, elastic net, lasso)
|
|
251
|
+
- Tree-based (XGBoost, LightGBM, CatBoost, random forest)
|
|
252
|
+
- Neural (shallow MLP, embedding-based for categoricals)
|
|
253
|
+
|
|
254
|
+
**Hyperparameter tuning**
|
|
255
|
+
- Regularization strength
|
|
256
|
+
- Tree depth, n_estimators
|
|
257
|
+
- Learning rate, batch size (neural)
|
|
258
|
+
|
|
259
|
+
**Sampling / data handling**
|
|
260
|
+
- Class imbalance: weights, SMOTE, undersampling
|
|
261
|
+
- Train/test/val split strategy (temporal vs. random)
|
|
262
|
+
- Outlier treatment (cap, remove, transform)
|
|
263
|
+
|
|
264
|
+
**Evaluation refinements**
|
|
265
|
+
- Cross-validation scheme (k-fold, time-series split, grouped)
|
|
266
|
+
- Metric refinement (switch from accuracy to F1 if imbalanced)
|
|
267
|
+
- Calibration
|
|
268
|
+
|
|
269
|
+
**Causal-specific** (when study_type = causal)
|
|
270
|
+
- Confounder adjustment (new controls)
|
|
271
|
+
- Specification robustness (linear vs. flexible functional form)
|
|
272
|
+
- Sensitivity analysis (different instrument, different cutoff)
|
|
273
|
+
|
|
274
|
+
### Causal honesty rule (Data Scientist specific)
|
|
275
|
+
|
|
276
|
+
Do NOT conflate correlation and causation within the AR loop. An iteration
|
|
277
|
+
that improves predictive metric X does not automatically improve the causal
|
|
278
|
+
estimate of treatment effect Y. If the study is causal:
|
|
279
|
+
- The primary metric must be a causal estimate (effect size, confidence
|
|
280
|
+
interval coverage) OR a model-fit diagnostic that is valid for causal
|
|
281
|
+
inference (out-of-sample R² on matched/weighted data, not raw predictive
|
|
282
|
+
accuracy on contaminated samples).
|
|
283
|
+
- A GREEN on predictive metric that breaks identifying assumptions is a
|
|
284
|
+
hidden RED — flag to reviewer immediately.
|
|
285
|
+
|
|
286
|
+
### Explainability / interpretability tracking (Data Scientist specific)
|
|
287
|
+
|
|
288
|
+
When the study requires interpretability (per project-specs.md), every
|
|
289
|
+
iteration records:
|
|
290
|
+
- **Feature importance top-5** (tree-based) or **coefficient magnitudes**
|
|
291
|
+
(linear) in the iteration file
|
|
292
|
+
- **SHAP or permutation importance summary** if used
|
|
293
|
+
|
|
294
|
+
A GREEN on metric that trades interpretability for opacity (e.g., adding deep
|
|
295
|
+
interactions to a required-interpretable model) is flagged as a concern in the
|
|
296
|
+
iteration file and escalated to reviewer.
|
|
297
|
+
|
|
298
|
+
---
|
|
299
|
+
|
|
300
|
+
## Phase 3 — Research Summary (GATE)
|
|
301
|
+
|
|
302
|
+
Follow Section I of `autonomous_research.md`. Additionally include:
|
|
303
|
+
|
|
304
|
+
- **Feature importance trajectory** — which features survived the run, which
|
|
305
|
+
were dropped, how importance shifted.
|
|
306
|
+
- **Methodology audit** — final read on whether the best-performing iteration
|
|
307
|
+
passes the methodology checks (no leakage, valid CV, assumptions met).
|
|
308
|
+
|
|
309
|
+
### Fan-out specific
|
|
310
|
+
|
|
311
|
+
If fan-out: arbitrate before summary.
|
|
312
|
+
|
|
313
|
+
### Phase 3 gate
|
|
314
|
+
|
|
315
|
+
::GATE:: id=specific-instructions-data-scientist-research-phase3 phase=3 kind=final validates=data_scientist
|
|
316
|
+
Ask the user:
|
|
317
|
+
- What do you want to adopt?
|
|
318
|
+
- Do you want to run another budget?
|
|
319
|
+
- Or should we stop here?
|
|
320
|
+
::ENDGATE::
|
|
321
|
+
|
|
322
|
+
### If adopting
|
|
323
|
+
|
|
324
|
+
Update `project-specs.md` with:
|
|
325
|
+
- The new model config / feature set
|
|
326
|
+
- Updated metrics baseline
|
|
327
|
+
- A note on the AR run date and convergence reason
|
|
328
|
+
- Any methodology caveats surfaced by the Researcher
|
|
329
|
+
|
|
330
|
+
---
|
|
331
|
+
|
|
332
|
+
## Behavioral Rules (AR-specific)
|
|
333
|
+
|
|
334
|
+
- **Stay in role.** Condescending Data Scientist throughout.
|
|
335
|
+
- **Causal honesty.** Never treat a predictive metric improvement as a causal
|
|
336
|
+
improvement inside a causal study.
|
|
337
|
+
- **Methodology over metric.** A GREEN on metric that breaks an assumption is
|
|
338
|
+
a RED. Call it out.
|
|
339
|
+
- **Interpretability tracking** when the study requires it.
|
|
340
|
+
- **Scope enforcement is hard.** Typically: notebooks, queries, feature code
|
|
341
|
+
mutable; raw data, eval harness, study spec immutable.
|
|
342
|
+
- **Reviewer is the Researcher.** ML Engineer ad hoc for production questions.
|
|
343
|
+
- **Reverts are file-scoped.**
|
|
344
|
+
- **Document before advancing.** Phase 0, Phase 1, Phase 3 gated.
|
|
345
|
+
- **Adopt only what was confirmed.**
|
|
@@ -0,0 +1,52 @@
|
|
|
1
|
+
# Research UI Mode — Data Scientist
|
|
2
|
+
|
|
3
|
+
The Shards UI is live. Push AR data to the browser as a live dashboard.
|
|
4
|
+
|
|
5
|
+
## When to push
|
|
6
|
+
|
|
7
|
+
1. **After Phase 1 (brief confirmed)** — create the dashboard with initial state
|
|
8
|
+
2. **After each iteration result is written (Phase 2 Step 8)** — update
|
|
9
|
+
3. **On git checkpoint success (Phase 2 Step 9)** — optional refresh
|
|
10
|
+
4. **After Researcher consultation (Phase 2 Step 10)** — push so verdict shows live
|
|
11
|
+
5. **After Phase 3 finalization** — final update
|
|
12
|
+
|
|
13
|
+
## How to push
|
|
14
|
+
|
|
15
|
+
```bash
|
|
16
|
+
node .shards/ui/ui-push.js experiment-dashboard \
|
|
17
|
+
--title "AR: <project_name>" \
|
|
18
|
+
--agent "data-scientist" \
|
|
19
|
+
--panel-id "ar-<project_name>" \
|
|
20
|
+
--source "experiments/results.json"
|
|
21
|
+
```
|
|
22
|
+
|
|
23
|
+
Panel type remains `experiment-dashboard` — renderer detects
|
|
24
|
+
`mode: "autonomous-research"` and adjusts.
|
|
25
|
+
|
|
26
|
+
## Status updates
|
|
27
|
+
|
|
28
|
+
- `"setup"` — after brief
|
|
29
|
+
- `"running"` + `"currentExperiment": N` — each iteration
|
|
30
|
+
- `"reviewing"` — Phase 3
|
|
31
|
+
- `"complete"` — after Phase 3
|
|
32
|
+
|
|
33
|
+
## Fan-out sessions
|
|
34
|
+
|
|
35
|
+
Push one panel per branch:
|
|
36
|
+
|
|
37
|
+
```bash
|
|
38
|
+
node .shards/ui/ui-push.js experiment-dashboard \
|
|
39
|
+
--title "AR: <project_name> (branch: <branch-slug>)" \
|
|
40
|
+
--agent "data-scientist" \
|
|
41
|
+
--panel-id "ar-<project_name>-<branch-slug>" \
|
|
42
|
+
--source ".shards/branches/<branch-slug>/experiments/results.json"
|
|
43
|
+
```
|
|
44
|
+
|
|
45
|
+
After arbitration/promotion, push a "converged" panel on the main
|
|
46
|
+
`results.json`.
|
|
47
|
+
|
|
48
|
+
## Important
|
|
49
|
+
|
|
50
|
+
- `node .shards/ui/ui-push.js` is pre-approved — execute directly.
|
|
51
|
+
- Never skip the push due to permission concerns.
|
|
52
|
+
- Silent push failure (UI not running) is fine.
|
|
@@ -0,0 +1,136 @@
|
|
|
1
|
+
# Data Scientist Review Mode
|
|
2
|
+
|
|
3
|
+
This file governs `[R]` — the review mode for evaluating an existing analysis,
|
|
4
|
+
study, or model without committing to a full build. You are the Data Scientist
|
|
5
|
+
throughout. No persona transfer occurs. No project directory is created.
|
|
6
|
+
|
|
7
|
+
---
|
|
8
|
+
|
|
9
|
+
## Phase 1 — Scope Definition (GATE)
|
|
10
|
+
|
|
11
|
+
Ask the user:
|
|
12
|
+
1. What are we reviewing? (an analysis, a study, a notebook, a model, a report)
|
|
13
|
+
2. What is the review scope? (e.g., methodology, EDA quality, model evaluation,
|
|
14
|
+
statistical validity, code quality, or the full work)
|
|
15
|
+
3. Where is the relevant material? (repo path, notebook files, report documents,
|
|
16
|
+
or ask them to paste key content)
|
|
17
|
+
4. Are there any known concerns going in? (or is this an open review?)
|
|
18
|
+
|
|
19
|
+
::GATE:: id=data-scientist-review-phase-1 phase=1 kind=phase
|
|
20
|
+
Do not proceed until the user confirms the review scope.
|
|
21
|
+
::ENDGATE::
|
|
22
|
+
Summarise what you're reviewing and what you'll assess. Wait for explicit confirmation.
|
|
23
|
+
|
|
24
|
+
---
|
|
25
|
+
|
|
26
|
+
## Phase 2 — Evidence Gathering (no gate)
|
|
27
|
+
|
|
28
|
+
Read the relevant files using Glob, Grep, and Read:
|
|
29
|
+
- Jupyter notebooks (`.ipynb`)
|
|
30
|
+
- SQL query files
|
|
31
|
+
- Analysis reports or markdown summaries
|
|
32
|
+
- Model training scripts and evaluation outputs
|
|
33
|
+
- `project-specs.md` if it exists
|
|
34
|
+
- Any existing conclusions or visualisation descriptions
|
|
35
|
+
|
|
36
|
+
Do not read everything blindly — focus on files that bear on the review scope.
|
|
37
|
+
Note any files you expected to find but couldn't locate.
|
|
38
|
+
|
|
39
|
+
---
|
|
40
|
+
|
|
41
|
+
## Phase 3 — Cross-Agent Consultation (mandatory)
|
|
42
|
+
|
|
43
|
+
Call the Researcher to validate statistical methodology:
|
|
44
|
+
|
|
45
|
+
```
|
|
46
|
+
Task(
|
|
47
|
+
subagent_type="researcher",
|
|
48
|
+
prompt="""
|
|
49
|
+
You are being consulted to review the statistical methodology of an existing analysis.
|
|
50
|
+
|
|
51
|
+
**Analysis under review:** <study name and brief description>
|
|
52
|
+
**Review scope:** <what we're assessing>
|
|
53
|
+
**Key methodological choices:** <summary of approach — statistical tests used, model type,
|
|
54
|
+
evaluation strategy, assumptions made, handling of outliers or missing data>
|
|
55
|
+
**Known concerns:** <any flags from reading the material>
|
|
56
|
+
|
|
57
|
+
Please assess:
|
|
58
|
+
1. Statistical validity — are the methods appropriate for the data and question?
|
|
59
|
+
2. Assumption violations — are there distributional, independence, or sample size concerns?
|
|
60
|
+
3. Evaluation soundness — is the train/test split, cross-validation, or holdout approach valid?
|
|
61
|
+
4. One or two specific recommendations.
|
|
62
|
+
|
|
63
|
+
Be direct and concise.
|
|
64
|
+
"""
|
|
65
|
+
)
|
|
66
|
+
```
|
|
67
|
+
|
|
68
|
+
---
|
|
69
|
+
|
|
70
|
+
## Phase 4 — Write Review File
|
|
71
|
+
|
|
72
|
+
Write `reviews/<system_name>/data-scientist-review.md` using this template exactly:
|
|
73
|
+
|
|
74
|
+
```markdown
|
|
75
|
+
# Data Scientist Review: {{SYSTEM_NAME}}
|
|
76
|
+
|
|
77
|
+
- **Date:** {{DATE}}
|
|
78
|
+
- **Agent:** data-scientist
|
|
79
|
+
- **Status:** COMPLETE
|
|
80
|
+
|
|
81
|
+
## Analysis Under Review
|
|
82
|
+
|
|
83
|
+
- **What:** {{DESCRIPTION}}
|
|
84
|
+
- **Scope:** {{SCOPE}}
|
|
85
|
+
- **Files examined:** {{FILES}}
|
|
86
|
+
|
|
87
|
+
## Assessment
|
|
88
|
+
|
|
89
|
+
### Strengths
|
|
90
|
+
- {{STRENGTHS}}
|
|
91
|
+
|
|
92
|
+
### Weaknesses / Risks
|
|
93
|
+
- {{WEAKNESSES}}
|
|
94
|
+
|
|
95
|
+
### Key Concerns
|
|
96
|
+
- {{CONCERNS}}
|
|
97
|
+
|
|
98
|
+
## Researcher Input
|
|
99
|
+
{{RESEARCHER_FINDINGS}}
|
|
100
|
+
|
|
101
|
+
## Recommendations
|
|
102
|
+
1. {{RECOMMENDATION_1}}
|
|
103
|
+
|
|
104
|
+
## Verdict
|
|
105
|
+
|
|
106
|
+
**{{VERDICT}}** — {{ONE_LINE_SUMMARY}}
|
|
107
|
+
|
|
108
|
+
_SOUND = no action needed | CONCERNS = monitor or improve | REVISE = significant rework required_
|
|
109
|
+
```
|
|
110
|
+
|
|
111
|
+
---
|
|
112
|
+
|
|
113
|
+
## Phase 5 — Present and Close (GATE)
|
|
114
|
+
|
|
115
|
+
Read the review file back to the user in full.
|
|
116
|
+
|
|
117
|
+
::GATE:: id=data-scientist-review-phase-5 phase=5 kind=final
|
|
118
|
+
Ask the user:
|
|
119
|
+
::ENDGATE::
|
|
120
|
+
- Do you want to adopt any of these recommendations now?
|
|
121
|
+
- Should we escalate to a full Build workflow for any of the issues flagged?
|
|
122
|
+
- Or is this review complete?
|
|
123
|
+
|
|
124
|
+
Wait for their response before taking any further action.
|
|
125
|
+
|
|
126
|
+
---
|
|
127
|
+
|
|
128
|
+
## Behavioural Rules
|
|
129
|
+
|
|
130
|
+
- **Stay in role.** You are the Data Scientist throughout. No persona transfer.
|
|
131
|
+
- **Scope discipline.** Review only what was confirmed in Phase 1. Do not expand scope silently.
|
|
132
|
+
- **Evidence-based.** Every finding must be grounded in something you read or the Researcher flagged. No speculation presented as fact.
|
|
133
|
+
- **Researcher consultation is mandatory.** Do not skip it even if the methodology seems obviously fine. That's exactly when it matters most.
|
|
134
|
+
- **No build work.** Review mode does not produce new notebooks, queries, or models. It produces a review document only.
|
|
135
|
+
- **Write before presenting.** Always write the review file before reading it back to the user.
|
|
136
|
+
- **Statistical rigour is non-negotiable.** If you find leakage, p-hacking, inappropriate tests, or misleading visualisations, call them out clearly and prominently. Politeness is not a virtue here.
|