@pikaa-ai/pikaa 0.3.23 → 0.3.25

This diff represents the content of publicly available package versions that have been released to one of the supported registries. The information contained in this diff is provided for informational purposes only and reflects changes between package versions as they appear in their respective public registries.
Files changed (191) hide show
  1. package/assets/brand/orbit-logo-option4-whale.jpg +0 -0
  2. package/assets/brand/orbit-logo.jpg +0 -0
  3. package/assets/brand/orbit-logo.png +0 -0
  4. package/assets/brand/orbit-logo.svg +3 -0
  5. package/dist/cli.js +407 -219
  6. package/dist/index.js +7 -2
  7. package/package.json +1 -2
  8. package/skills/adaptyv/SKILL.md +0 -240
  9. package/skills/aeon/SKILL.md +0 -402
  10. package/skills/analytical-method-validation/SKILL.md +0 -299
  11. package/skills/anndata/SKILL.md +0 -431
  12. package/skills/arbor/SKILL.md +0 -152
  13. package/skills/arboreto/SKILL.md +0 -267
  14. package/skills/astropy/SKILL.md +0 -353
  15. package/skills/autoskill/SKILL.md +0 -233
  16. package/skills/benchling-integration/SKILL.md +0 -229
  17. package/skills/bgpt-paper-search/SKILL.md +0 -75
  18. package/skills/bids/SKILL.md +0 -237
  19. package/skills/biopython/SKILL.md +0 -472
  20. package/skills/bioservices/SKILL.md +0 -399
  21. package/skills/bulk-rnaseq/SKILL.md +0 -198
  22. package/skills/cellxgene-census/SKILL.md +0 -283
  23. package/skills/cirq/SKILL.md +0 -370
  24. package/skills/citation-management/SKILL.md +0 -329
  25. package/skills/clinical-decision-support/SKILL.md +0 -238
  26. package/skills/clinical-decision-support/references/README.md +0 -62
  27. package/skills/clinical-reports/SKILL.md +0 -248
  28. package/skills/clinical-reports/references/README.md +0 -34
  29. package/skills/cobrapy/SKILL.md +0 -496
  30. package/skills/consciousness-council/SKILL.md +0 -151
  31. package/skills/dask/SKILL.md +0 -482
  32. package/skills/database-lookup/SKILL.md +0 -386
  33. package/skills/datamol/SKILL.md +0 -200
  34. package/skills/deepchem/SKILL.md +0 -244
  35. package/skills/deepspot-m/SKILL.md +0 -175
  36. package/skills/deeptools/SKILL.md +0 -412
  37. package/skills/depmap/SKILL.md +0 -301
  38. package/skills/dhdna-profiler/SKILL.md +0 -184
  39. package/skills/diffdock/SKILL.md +0 -488
  40. package/skills/dnanexus-integration/SKILL.md +0 -325
  41. package/skills/docx/SKILL.md +0 -99
  42. package/skills/esm/SKILL.md +0 -334
  43. package/skills/etetoolkit/SKILL.md +0 -327
  44. package/skills/exa-search/SKILL.md +0 -102
  45. package/skills/executing-plans/SKILL.md +0 -14
  46. package/skills/experimental-design/SKILL.md +0 -234
  47. package/skills/exploratory-data-analysis/SKILL.md +0 -280
  48. package/skills/flowio/SKILL.md +0 -310
  49. package/skills/fluidsim/SKILL.md +0 -279
  50. package/skills/frontend-design/SKILL.md +0 -100
  51. package/skills/generate-image/SKILL.md +0 -304
  52. package/skills/geniml/SKILL.md +0 -310
  53. package/skills/genomic-coordinates/SKILL.md +0 -189
  54. package/skills/genomic-intelligence/SKILL.md +0 -243
  55. package/skills/geomaster/README.md +0 -105
  56. package/skills/geomaster/SKILL.md +0 -366
  57. package/skills/geopandas/SKILL.md +0 -250
  58. package/skills/get-available-resources/SKILL.md +0 -260
  59. package/skills/gget/SKILL.md +0 -153
  60. package/skills/ginkgo-cloud-lab/SKILL.md +0 -106
  61. package/skills/glycoengineering/SKILL.md +0 -339
  62. package/skills/gtars/SKILL.md +0 -282
  63. package/skills/guardian-rails/SKILL.md +0 -54
  64. package/skills/histolab/SKILL.md +0 -243
  65. package/skills/hugging-science/SKILL.md +0 -132
  66. package/skills/hypogenic/SKILL.md +0 -290
  67. package/skills/hypothesis-generation/SKILL.md +0 -264
  68. package/skills/imaging-data-commons/SKILL.md +0 -496
  69. package/skills/infographics/SKILL.md +0 -315
  70. package/skills/iso-standards-readiness/SKILL.md +0 -352
  71. package/skills/lab-hardware-cad/SKILL.md +0 -372
  72. package/skills/labarchive-integration/SKILL.md +0 -216
  73. package/skills/lamindb/SKILL.md +0 -408
  74. package/skills/latchbio-integration/SKILL.md +0 -227
  75. package/skills/latex-posters/SKILL.md +0 -369
  76. package/skills/latex-posters/references/README.md +0 -439
  77. package/skills/liteparse/SKILL.md +0 -295
  78. package/skills/literature-review/SKILL.md +0 -263
  79. package/skills/markdown-mermaid-writing/SKILL.md +0 -322
  80. package/skills/market-research-reports/SKILL.md +0 -337
  81. package/skills/markitdown/SKILL.md +0 -264
  82. package/skills/matchms/SKILL.md +0 -276
  83. package/skills/matlab/SKILL.md +0 -274
  84. package/skills/matplotlib/SKILL.md +0 -378
  85. package/skills/medchem/SKILL.md +0 -321
  86. package/skills/modal/SKILL.md +0 -468
  87. package/skills/molecular-dynamics/SKILL.md +0 -458
  88. package/skills/molfeat/SKILL.md +0 -348
  89. package/skills/ncats-arax/SKILL.md +0 -178
  90. package/skills/networkx/SKILL.md +0 -440
  91. package/skills/neurokit2/SKILL.md +0 -323
  92. package/skills/neuropixels-analysis/SKILL.md +0 -412
  93. package/skills/nextflow/SKILL.md +0 -195
  94. package/skills/omero-integration/SKILL.md +0 -222
  95. package/skills/onekgpd/SKILL.md +0 -371
  96. package/skills/ontology-term-resolution/SKILL.md +0 -147
  97. package/skills/open-notebook/SKILL.md +0 -297
  98. package/skills/openpiv/SKILL.md +0 -469
  99. package/skills/opentrons-integration/SKILL.md +0 -322
  100. package/skills/optimize-for-gpu/SKILL.md +0 -176
  101. package/skills/owasp-top10/SKILL.md +0 -48
  102. package/skills/pacsomatic/LICENSE +0 -21
  103. package/skills/pacsomatic/SKILL.md +0 -150
  104. package/skills/paper-lookup/SKILL.md +0 -263
  105. package/skills/paperclip/SKILL.md +0 -413
  106. package/skills/paperzilla/SKILL.md +0 -159
  107. package/skills/parallel-web/SKILL.md +0 -128
  108. package/skills/pathml/SKILL.md +0 -222
  109. package/skills/pathogen-variant-surveillance/SKILL.md +0 -208
  110. package/skills/pathway-enrichment/SKILL.md +0 -194
  111. package/skills/pdf/SKILL.md +0 -322
  112. package/skills/peer-review/SKILL.md +0 -288
  113. package/skills/penetration-testing/SKILL.md +0 -31
  114. package/skills/pennylane/SKILL.md +0 -240
  115. package/skills/phylogenetics/SKILL.md +0 -409
  116. package/skills/pi-agent/SKILL.md +0 -83
  117. package/skills/pkpd-modeling/SKILL.md +0 -381
  118. package/skills/polars/SKILL.md +0 -393
  119. package/skills/polars-bio/SKILL.md +0 -379
  120. package/skills/ponytail/SKILL.md +0 -31
  121. package/skills/ponytail-audit/SKILL.md +0 -18
  122. package/skills/pptx/SKILL.md +0 -246
  123. package/skills/pptx-posters/SKILL.md +0 -258
  124. package/skills/primekg/SKILL.md +0 -99
  125. package/skills/protocolsio-integration/SKILL.md +0 -236
  126. package/skills/pufferlib/SKILL.md +0 -328
  127. package/skills/pydeseq2/SKILL.md +0 -369
  128. package/skills/pydicom/SKILL.md +0 -381
  129. package/skills/pyhealth/SKILL.md +0 -124
  130. package/skills/pylabrobot/SKILL.md +0 -216
  131. package/skills/pymatgen/SKILL.md +0 -404
  132. package/skills/pymc/SKILL.md +0 -310
  133. package/skills/pymoo/SKILL.md +0 -276
  134. package/skills/pyopenms/SKILL.md +0 -179
  135. package/skills/pysam/SKILL.md +0 -330
  136. package/skills/pytdc/SKILL.md +0 -297
  137. package/skills/pytorch-lightning/SKILL.md +0 -191
  138. package/skills/pyzotero/SKILL.md +0 -137
  139. package/skills/qiskit/SKILL.md +0 -259
  140. package/skills/qutip/SKILL.md +0 -317
  141. package/skills/rdkit/SKILL.md +0 -94
  142. package/skills/relsa-severity-assessment/SKILL.md +0 -354
  143. package/skills/research-grants/SKILL.md +0 -296
  144. package/skills/research-grants/references/README.md +0 -287
  145. package/skills/research-lookup/README.md +0 -106
  146. package/skills/research-lookup/SKILL.md +0 -338
  147. package/skills/rowan/SKILL.md +0 -398
  148. package/skills/scanpy/SKILL.md +0 -303
  149. package/skills/scholar-evaluation/SKILL.md +0 -296
  150. package/skills/scientific-brainstorming/SKILL.md +0 -282
  151. package/skills/scientific-critical-thinking/SKILL.md +0 -180
  152. package/skills/scientific-schematics/SKILL.md +0 -370
  153. package/skills/scientific-slides/SKILL.md +0 -379
  154. package/skills/scientific-visualization/SKILL.md +0 -285
  155. package/skills/scientific-writing/SKILL.md +0 -356
  156. package/skills/scikit-bio/SKILL.md +0 -470
  157. package/skills/scikit-learn/SKILL.md +0 -324
  158. package/skills/scikit-survival/SKILL.md +0 -313
  159. package/skills/scvelo/SKILL.md +0 -328
  160. package/skills/scvi-tools/SKILL.md +0 -201
  161. package/skills/seaborn/SKILL.md +0 -254
  162. package/skills/security-auditor/SKILL.md +0 -37
  163. package/skills/shap/SKILL.md +0 -282
  164. package/skills/simpy/SKILL.md +0 -283
  165. package/skills/stable-baselines3/SKILL.md +0 -325
  166. package/skills/statistical-analysis/SKILL.md +0 -446
  167. package/skills/statistical-power/SKILL.md +0 -200
  168. package/skills/statsmodels/SKILL.md +0 -238
  169. package/skills/sympy/SKILL.md +0 -354
  170. package/skills/systematic-debugging/SKILL.md +0 -35
  171. package/skills/tamarind/SKILL.md +0 -285
  172. package/skills/tdd/SKILL.md +0 -26
  173. package/skills/tiledbvcf/SKILL.md +0 -456
  174. package/skills/timesfm-forecasting/SKILL.md +0 -408
  175. package/skills/timesfm-forecasting/examples/global-temperature/README.md +0 -178
  176. package/skills/torch-geometric/SKILL.md +0 -458
  177. package/skills/torchdrug/SKILL.md +0 -241
  178. package/skills/transformers/SKILL.md +0 -195
  179. package/skills/treatment-plans/SKILL.md +0 -174
  180. package/skills/treatment-plans/references/README.md +0 -19
  181. package/skills/umap-learn/SKILL.md +0 -488
  182. package/skills/uncertainty-and-units/SKILL.md +0 -384
  183. package/skills/usfiscaldata/SKILL.md +0 -171
  184. package/skills/vaex/SKILL.md +0 -204
  185. package/skills/venue-templates/SKILL.md +0 -269
  186. package/skills/verification-before-completion/SKILL.md +0 -22
  187. package/skills/waypoint-bio/SKILL.md +0 -273
  188. package/skills/what-if-oracle/SKILL.md +0 -184
  189. package/skills/writing-plans/SKILL.md +0 -15
  190. package/skills/xlsx/SKILL.md +0 -110
  191. package/skills/zarr-python/SKILL.md +0 -241
@@ -1,282 +0,0 @@
1
- ---
2
- name: scientific-brainstorming
3
- description: Facilitates evidence-aware scientific ideation with independent generation, structured discussion, explicit assumptions, transparent evaluation, adversarial review, and decision logs. Use for early-stage research brainstorming or prioritizing candidate directions; hand off empirical validation, study design, ethics or regulatory review, and clinical questions to appropriate experts or skills.
4
- license: MIT
5
- compatibility: Core guidance works in any Agent Skills-compatible host. Optional bundled CLIs require Python 3.11+ and use only the standard library; they make no network or LLM calls and require no credentials.
6
- metadata:
7
- version: "1.1"
8
- skill-author: "K-Dense Inc."
9
- ---
10
-
11
- # Scientific Brainstorming
12
-
13
- ## Purpose and boundaries
14
-
15
- Use this skill to create, organize, challenge, and transparently prioritize
16
- candidate research directions. Treat every output as a **proposal**, not a
17
- finding. Creativity methods can alter participation and idea yield, but no
18
- method universally improves originality, usefulness, or scientific validity.
19
- The evidence base and its limits are summarized in
20
- `references/sources.md`.
21
-
22
- Keep these activities separate:
23
-
24
- - **Ideation** creates questions, mechanisms, alternatives, or study concepts.
25
- - **Evidence assessment** checks what reliable literature and data support.
26
- - **Hypothesis validation** requires observations, predictions, suitable
27
- designs, analyses, and independent scrutiny; brainstorming cannot validate a
28
- hypothesis.
29
- - **Ethics, biosafety, dual-use, regulatory, and institutional review** require
30
- the relevant authorized reviewers. A brainstorm is never approval.
31
- - **Clinical advice** requires qualified clinicians and patient-specific
32
- context. Do not turn research ideas into diagnosis or treatment guidance.
33
-
34
- For an observation-led testable hypothesis, hand off to
35
- `hypothesis-generation`. For study architecture, use `experimental-design`;
36
- for sample size, `statistical-power`; for existing evidence,
37
- `literature-review`; and for analysis, `statistical-analysis`.
38
-
39
- ## Operating rules
40
-
41
- 1. Label claims as **idea**, **assumption**, **prediction**, **located
42
- evidence**, or **decision**. Never blur these categories.
43
- 2. Generate independently before exposing participants to other people's or
44
- AI-generated ideas. Face-to-face turn-taking can block production, and
45
- examples can anchor later output.
46
- 3. Preserve minority views, negative evidence, uncertainty, and abstentions.
47
- Consensus is not truth and vote counts are not effect sizes.
48
- 4. Record provenance without exposing confidential, personal, controlled, or
49
- unpublished information.
50
- 5. Define evaluation criteria and directions before scoring. Keep raw ratings,
51
- reasons, ranges, and disagreement visible.
52
- 6. Search the literature **after an initial independent round** when practical,
53
- then deliberately reopen ideation. This reduces early anchoring without
54
- mistaking an incomplete search for a research gap.
55
- 7. Do not automatically select a “winner.” Scores are traceable decision aids;
56
- qualitative judgment, uncertainty, feasibility, and ethics gates remain
57
- controlling.
58
-
59
- ## Reproducible workflow
60
-
61
- ### 1. Scope the session
62
-
63
- Write one focal question and record:
64
-
65
- - purpose, audience, decision owner, and time horizon;
66
- - in-scope and out-of-scope topics;
67
- - constraints that are real, assumed, negotiable, or unknown;
68
- - current knowledge, unresolved observations, and prohibited outputs;
69
- - whether human participants, animals, clinical care, sensitive data,
70
- pathogens, controlled technologies, or environmental release could be
71
- implicated.
72
-
73
- If the request seeks patient-specific care, evasion of oversight, harmful
74
- optimization, or operationally enabling dual-use details, stop ideation and
75
- route to the appropriate professional or institutional process.
76
-
77
- ### 2. Diversify perspectives deliberately
78
-
79
- Invite relevant methodological, domain, implementation, statistical, safety,
80
- ethics, stakeholder, and lived-experience perspectives. Diversity is not a
81
- guarantee of creativity: explain whose perspective is represented, missing, or
82
- structurally disadvantaged. Use accessible participation modes and
83
- pseudonymous participant IDs where appropriate.
84
-
85
- The facilitator should disclose conflicts, avoid offering a preferred answer
86
- first, prevent senior members from dominating, and ask leaders to contribute
87
- after the independent round.
88
-
89
- ### 3. Generate independently
90
-
91
- Give everyone the same neutral prompt, constraints, and fixed time window.
92
- Participants write ideas privately and in parallel before discussion. For each
93
- idea, capture:
94
-
95
- - a stable ID and one-sentence statement;
96
- - contributor ID(s) and stage (`independent`, `discussion`, or `post-check`);
97
- - origin (`human`, `AI-assisted`, `literature-inspired`, `mixed`, or `other`);
98
- - assumptions, predicted observations, uncertainties, and possible
99
- disconfirming evidence;
100
- - source identifiers for literature-inspired ideas and tool/purpose disclosure
101
- for AI assistance.
102
-
103
- Do not show example solutions before this round unless examples are necessary;
104
- if they are, record them as potential anchors.
105
-
106
- ### 4. Share without immediate evaluation
107
-
108
- Use round-robin or pooled silent sharing. Clarify wording without advocacy.
109
- Permit a private or anonymous channel. Ask each participant what is missing,
110
- what contradicts the dominant framing, and which idea became less obvious
111
- after hearing the group.
112
-
113
- ### 5. Cluster structurally
114
-
115
- Group ideas by an explicit relation such as shared outcome, mechanism,
116
- population, scale, or method. Keep original IDs and text. Record merges and
117
- splits. Similar wording is not proof of semantic equivalence; retain distinct
118
- ideas when their assumptions, intervention, population, or predictions differ.
119
- See `references/facilitation_workflows.md`.
120
-
121
- ### 6. Define transparent criteria
122
-
123
- Before rating, define each criterion, direction, scale anchors, evidence
124
- needed, conflicts, and explicit weights. Common dimensions include:
125
-
126
- - potential information gain and discriminating predictions;
127
- - relevance to the scoped question;
128
- - originality relative to the checked literature, not merely to the room;
129
- - feasibility, resources, and reversibility;
130
- - methodological rigor and vulnerability to bias;
131
- - ethics, safety, equity, dual-use, and regulatory burden;
132
- - value if the result is null or contradicts the favored mechanism.
133
-
134
- Use ranges or confidence labels where assessors are uncertain. Do not hide
135
- vetoes inside an averaged score. See `references/idea_evaluation.md`.
136
-
137
- ### 7. Run adversarial review
138
-
139
- Assign a reviewer who did not originate each shortlisted idea. Ask:
140
-
141
- - What observation would make this idea wrong or uninformative?
142
- - Which alternative explanation fits the same predicted result?
143
- - What hidden dependency, measurement failure, confounder, or selection effect
144
- could dominate?
145
- - Are authority, anchoring, group loyalty, publication incentives, or an
146
- attractive technology driving preference?
147
- - Could this cause harm, worsen inequity, expose sensitive information, or
148
- enable misuse?
149
-
150
- Record the response, mitigation, residual uncertainty, and whether the idea was
151
- revised—not just pass/fail.
152
-
153
- ### 8. Check literature and evidence
154
-
155
- Search authoritative databases, primary studies, methods guidance, negative
156
- results, and adjacent fields. Verify every citation at its source. For each
157
- idea, record query/date, sources screened, evidence for and against, and search
158
- limits. Use statuses such as `not-checked`, `search-incomplete`,
159
- `support-located`, `challenge-located`, or `mixed`.
160
-
161
- Absence from a bounded search does not establish novelty, and supportive
162
- literature does not validate a new mechanism. Reopen one short independent
163
- generation round after the evidence check.
164
-
165
- ### 9. Apply feasibility, rigor, and ethics gates
166
-
167
- Before advancing an idea, identify the appropriate domain review:
168
-
169
- - For biomedical work, consider rigor of prior research, robust design,
170
- relevant biological variables, and resource authentication. When NIH policy
171
- applies, sex as a biological variable should be considered from the research
172
- question through design, analysis, and reporting; justify a single-sex scope
173
- with relevant evidence.
174
- - Route human-subjects, animal, biosafety, data-governance, export-control,
175
- clinical, environmental, and other regulated work to the relevant office.
176
- - Screen life-science and enabling-technology ideas for dual-use or misuse
177
- potential early. Current U.S. oversight is evolving; consult the institution
178
- and current agency policy rather than relying on a static checklist.
179
- - Do not upload sensitive, unpublished, proprietary, controlled, or personal
180
- information to an external AI service.
181
-
182
- An ethics or feasibility concern may require redesign, controlled handling, or
183
- stopping. A high creativity score never overrides a gate.
184
-
185
- ### 10. Decide and log
186
-
187
- The accountable human decision owner records:
188
-
189
- - candidates considered and criteria/weights used;
190
- - raw ratings, uncertainty ranges, dissent, abstentions, and sensitivity
191
- results;
192
- - literature and review dates;
193
- - gate outcomes and required approvals;
194
- - decision, rationale, rejected alternatives, unresolved risks, owner, and
195
- revisit trigger.
196
-
197
- Label the next action correctly: further search, consultation, simulation,
198
- pilot design, protocol development, preregistration, or no action. If a
199
- confirmatory study is planned, preregister hypotheses and analysis decisions
200
- before outcomes are known; report later deviations and exploratory work
201
- transparently. Preregistration improves transparency but is not peer review,
202
- ethical approval, or proof of validity.
203
-
204
- ## Bias and failure controls
205
-
206
- - **Production blocking:** private parallel generation before oral discussion.
207
- - **Anchoring and design fixation:** no leader answer or AI examples until the
208
- independent round; reopen generation after evidence review.
209
- - **Authority and status effects:** leader-last sharing, anonymous input,
210
- independent ratings, and visible dissent.
211
- - **Groupthink:** assign a genuine alternative-generation role, invite outside
212
- review, and document rejected options. Treat “groupthink” as a family of
213
- risks, not a single universally established diagnosis.
214
- - **Evaluation apprehension:** separate contribution from attribution where
215
- possible; critique ideas, not contributors.
216
- - **Premature convergence:** fixed divergence window followed by an explicit
217
- transition and predeclared criteria.
218
- - **False precision:** use anchored scales, uncertainty ranges, sensitivity
219
- analysis, and narrative review.
220
- - **Research-gap inflation:** record search boundaries and use “no direct
221
- evidence located,” not “never studied.”
222
- - **AI hallucination or homogenization:** human-first ideation, provenance,
223
- independent verification, multiple non-AI perspectives, and comparison for
224
- suspiciously repeated frames. See `references/responsible_ai.md`.
225
-
226
- ## Optional local CLIs
227
-
228
- The scripts are deterministic, standard-library utilities. They do not call a
229
- network service, LLM, or scientific database and do not make scientific
230
- conclusions.
231
-
232
- ```bash
233
- python scripts/session_scaffold.py --help
234
- python scripts/validate_register.py --help
235
- python scripts/evaluate_matrix.py --help
236
- ```
237
-
238
- Create a session register:
239
-
240
- ```bash
241
- python scripts/session_scaffold.py \
242
- --session-id "microbiome-01" \
243
- --title "Microbiome mechanism ideation" \
244
- --question "Which mechanisms could explain the scoped observation?" \
245
- --participant P01 --participant P02 \
246
- --output session.json
247
- ```
248
-
249
- Validate structure and provenance:
250
-
251
- ```bash
252
- python scripts/validate_register.py session.json --output validation.json
253
- ```
254
-
255
- Calculate a fully disclosed weighted matrix from CSV, including score intervals
256
- and one-at-a-time weight sensitivity:
257
-
258
- ```bash
259
- python scripts/evaluate_matrix.py scores.csv \
260
- --config criteria.json \
261
- --weight-delta 0.10 \
262
- --output matrix.json
263
- ```
264
-
265
- Outputs refuse symlinks and existing files unless `--force` is explicit; inputs
266
- and collection sizes are bounded. The validator checks structure, not truth.
267
- The matrix preserves qualitative review and uncertainty and leaves
268
- `decision` null. Input formats and interpretation are documented in
269
- `references/idea_evaluation.md`.
270
-
271
- ## Reference index
272
-
273
- - `references/brainstorming_methods.md` — evidence-calibrated method selection,
274
- nominal groups, Delphi, structured elicitation, and creative prompts.
275
- - `references/facilitation_workflows.md` — ready-to-run individual, group, and
276
- asynchronous session protocols plus provenance templates.
277
- - `references/idea_evaluation.md` — criteria, scoring formula, uncertainty,
278
- sensitivity analysis, gates, and decision logs.
279
- - `references/responsible_ai.md` — accountable AI assistance, confidentiality,
280
- hallucination, homogenization, disclosure, dual-use, and integrity.
281
- - `references/sources.md` — dated primary studies and official guidance
282
- consulted for this version.
@@ -1,180 +0,0 @@
1
- ---
2
- name: scientific-critical-thinking
3
- description: Evaluate scientific claims and evidence quality. Use for assessing experimental design validity, identifying biases and confounders, applying evidence grading frameworks (GRADE, Cochrane Risk of Bias), or teaching critical analysis. Best for understanding evidence quality, identifying flaws. For formal peer review writing use peer-review.
4
- allowed-tools: Read Write Edit
5
- license: MIT license
6
- compatibility: Analytical guidance needs no network. Optional figures via the scientific-schematics skill require OPENROUTER_API_KEY and outbound API access to OpenRouter.
7
- metadata:
8
- version: "1.2"
9
- skill-author: K-Dense Inc.
10
- ---
11
-
12
- # Scientific Critical Thinking
13
-
14
- ## Overview
15
-
16
- Critical thinking is a systematic process for evaluating scientific rigor. Assess methodology, experimental design, statistical validity, biases, confounding, and evidence quality using GRADE and Cochrane ROB frameworks. Apply this skill for critical analysis of scientific claims.
17
-
18
- ## When to Use This Skill
19
-
20
- This skill should be used when:
21
- - Evaluating research methodology and experimental design
22
- - Assessing statistical validity and evidence quality
23
- - Identifying biases and confounding in studies
24
- - Reviewing scientific claims and conclusions
25
- - Conducting systematic reviews or meta-analyses
26
- - Applying GRADE or Cochrane risk of bias assessments
27
- - Providing critical analysis of research papers
28
-
29
- ## Visual Aids (Optional)
30
-
31
- Only add figures when the **user explicitly requests** a diagram (for example, a GRADE flowchart, bias decision tree, or evidence-quality framework).
32
-
33
- **When figures help:**
34
- - Critical thinking framework diagrams
35
- - Bias identification decision trees
36
- - Evidence quality assessment flowcharts
37
- - GRADE or risk-of-bias evaluation frameworks
38
-
39
- **How to create figures:**
40
- - **Preferred:** Use the **scientific-schematics** skill for AI-generated diagrams from a natural-language description
41
- - **Alternative:** Build figures in your usual tools (draw.io, PowerPoint, matplotlib, etc.)
42
-
43
- From the `scientific-schematics` skill directory, with `OPENROUTER_API_KEY` set:
44
-
45
- ```bash
46
- python scripts/generate_schematic.py "GRADE evidence assessment flowchart with downgrade and upgrade factors" -o figures/grade_flowchart.png --doc-type report
47
- ```
48
-
49
- **Disclosure:** AI schematic generation sends your prompt to [OpenRouter](https://openrouter.ai/) (a third-party API). Do not include unpublished sensitive details unless that transmission is appropriate for your project.
50
-
51
- ---
52
-
53
- ## Core Capabilities
54
-
55
- Seven capability areas, each with the questions to ask and what the answers imply, are in
56
- [references/core_capabilities.md](references/core_capabilities.md):
57
-
58
- 1. **Methodology critique** — design, controls, confounding, and whether the method can
59
- answer the question asked.
60
- 2. **Bias detection** — selection, measurement, publication, and cognitive biases.
61
- 3. **Statistical analysis evaluation** — power, multiplicity, p-value misuse, effect sizes.
62
- 4. **Evidence quality assessment** — study hierarchy, replication, and strength of inference.
63
- 5. **Logical fallacy identification** — the fallacies that recur in scientific argument.
64
- 6. **Research design guidance** — how to strengthen a design before data collection.
65
- 7. **Claim evaluation** — separating what was shown from what is being asserted.
66
-
67
- Per-topic detail is in [references/scientific_method.md](references/scientific_method.md),
68
- [references/common_biases.md](references/common_biases.md),
69
- [references/statistical_pitfalls.md](references/statistical_pitfalls.md),
70
- [references/evidence_hierarchy.md](references/evidence_hierarchy.md),
71
- [references/logical_fallacies.md](references/logical_fallacies.md), and
72
- [references/experimental_design.md](references/experimental_design.md).
73
-
74
- ## Application Guidelines
75
-
76
- ### General Approach
77
-
78
- 1. **Be Constructive**
79
- - Identify strengths as well as weaknesses
80
- - Suggest improvements rather than just criticizing
81
- - Distinguish between fatal flaws and minor limitations
82
- - Recognize that all research has limitations
83
-
84
- 2. **Be Specific**
85
- - Point to specific instances (e.g., "Table 2 shows..." or "In the Methods section...")
86
- - Quote problematic statements
87
- - Provide concrete examples of issues
88
- - Reference specific principles or standards violated
89
-
90
- 3. **Be Proportionate**
91
- - Match criticism severity to issue importance
92
- - Distinguish between major threats to validity and minor concerns
93
- - Consider whether issues affect primary conclusions
94
- - Acknowledge uncertainty in your own assessments
95
-
96
- 4. **Apply Consistent Standards**
97
- - Use same criteria across all studies
98
- - Don't apply stricter standards to findings you dislike
99
- - Acknowledge your own potential biases
100
- - Base judgments on methodology, not results
101
-
102
- 5. **Consider Context**
103
- - Acknowledge practical and ethical constraints
104
- - Consider field-specific norms for effect sizes and methods
105
- - Recognize exploratory vs. confirmatory contexts
106
- - Account for resource limitations in evaluating studies
107
-
108
- ### When Providing Critique
109
-
110
- **Structure feedback as:**
111
-
112
- 1. **Summary:** Brief overview of what was evaluated
113
- 2. **Strengths:** What was done well (important for credibility and learning)
114
- 3. **Concerns:** Issues organized by severity
115
- - Critical issues (threaten validity of main conclusions)
116
- - Important issues (affect interpretation but not fatally)
117
- - Minor issues (worth noting but don't change conclusions)
118
- 4. **Specific Recommendations:** Actionable suggestions for improvement
119
- 5. **Overall Assessment:** Balanced conclusion about evidence quality and what can be concluded
120
-
121
- **Use precise terminology:**
122
- - Name specific biases, fallacies, and methodological issues
123
- - Reference established standards and guidelines
124
- - Cite principles from scientific methodology
125
- - Use technical terms accurately
126
-
127
- ### When Uncertain
128
-
129
- - **Acknowledge uncertainty:** "This could be X or Y; additional information needed is Z"
130
- - **Ask clarifying questions:** "Was [methodological detail] done? This affects interpretation."
131
- - **Provide conditional assessments:** "If X was done, then Y follows; if not, then Z is concern"
132
- - **Note what additional information would resolve uncertainty**
133
-
134
- ## Reference Materials
135
-
136
- This skill includes comprehensive reference materials that provide detailed frameworks for critical evaluation:
137
-
138
- - **`references/scientific_method.md`** - Core principles of scientific methodology, the scientific process, critical evaluation criteria, red flags in scientific claims, causal inference standards, peer review, and open science principles
139
-
140
- - **`references/common_biases.md`** - Comprehensive taxonomy of cognitive, experimental, methodological, statistical, and analysis biases with detection and mitigation strategies
141
-
142
- - **`references/statistical_pitfalls.md`** - Common statistical errors and misinterpretations including p-value misunderstandings, multiple comparisons problems, sample size issues, effect size mistakes, correlation/causation confusion, regression pitfalls, and meta-analysis issues
143
-
144
- - **`references/evidence_hierarchy.md`** - Traditional evidence hierarchy, GRADE system, study quality assessment criteria, domain-specific considerations, evidence synthesis principles, and practical decision frameworks
145
-
146
- - **`references/logical_fallacies.md`** - Logical fallacies common in scientific discourse organized by type (causation, generalization, authority, relevance, structure, statistical) with examples and detection strategies
147
-
148
- - **`references/experimental_design.md`** - Comprehensive experimental design checklist covering research questions, hypotheses, study design selection, variables, sampling, blinding, randomization, control groups, procedures, measurement, bias minimization, data management, statistical planning, ethical considerations, validity threats, and reporting standards
149
-
150
- **When to consult references:**
151
- - Load references into context when detailed frameworks are needed
152
- - Use grep to search references for specific topics: `grep -r "pattern" references/`
153
- - References provide depth; SKILL.md provides procedural guidance
154
- - Consult references for comprehensive lists, detailed criteria, and specific examples
155
-
156
- ## Remember
157
-
158
- **Scientific critical thinking is about:**
159
- - Systematic evaluation using established principles
160
- - Constructive critique that improves science
161
- - Proportional confidence to evidence strength
162
- - Transparency about uncertainty and limitations
163
- - Consistent application of standards
164
- - Recognition that all research has limitations
165
- - Balance between skepticism and openness to evidence
166
-
167
- **Always distinguish between:**
168
- - Data (what was observed) and interpretation (what it means)
169
- - Correlation and causation
170
- - Statistical significance and practical importance
171
- - Exploratory and confirmatory findings
172
- - What is known and what is uncertain
173
- - Evidence against a claim and evidence for the null
174
-
175
- **Goals of critical thinking:**
176
- 1. Identify strengths and weaknesses accurately
177
- 2. Determine what conclusions are supported
178
- 3. Recognize limitations and uncertainties
179
- 4. Suggest improvements for future work
180
- 5. Advance scientific understanding