Topic · Research & Science

Best experimental design skills for Claude Code, Codex and other agents.

Skills that design experiments and studies with sound controls and power.
skills
286
official
2

Experimental design skills, ranked

Ranked by score. Sort bymost stars,trending,newest,recently updated

Experimental design skills, ranked
#SkillRepositoryStarsUsed inTokensAuto-checkLicenceUpdated
1

Evaluate research rigor. An agent skill from weapp-tailwindcss/weapp-tailwindcss.

weapp-tailwindcss/weapp-tailwindcss1.9k23 repos~5.9kAutomated safety check: NotesMITtoday
2

Guided statistical analysis for research data - test selection, assumption checking, effect sizes, power analysis, Bayesian alternatives, and APA-formatted reporting.

spacering-net/codeg3.8k3 repos~5kAutomated safety check: PassMITtoday
3

Turns a refined research proposal into a claim-to-evidence-to-run-order roadmap instead of a sprawling benchmark wishlist.

zjYao36/Auto-Research-Refine1287 repos~2.3kAutomated safety check: NotesNo licence6 mo ago
4

Structures benchmark and evaluation papers around five pillars, with a completeness audit, an Introduction logic chain, a section skeleton and a pre-submission checklist.

HKUSTDial/Supervisor-Skills8.4k—~2.8kAutomated safety check: PassCC-BY-4.01 mo ago
5

Sample-size and statistical power calculations for planning studies.

spacering-net/codeg3.8k1 repo~3.6kAutomated safety check: NotesMITtoday
6

Chains research-refine and experiment-plan to turn a vague research direction into a focused proposal and a claim-driven experiment roadmap.

zjYao36/Auto-Research-Refine1286 repos~1.4kAutomated safety check: NotesNo licence6 mo ago
7

Builds a scoring-aligned outline for a mathematical modeling paper and a model selection plan with baseline, improvement and validation experiments.

yushui2022/MathModel-Skill4521 repo~1.8kAutomated safety check: PassMIT10 days ago
8

Turns a broad metabolic modelling topic into a concrete, paper-shaped plan with organism, model, perturbations, metrics and figures before any FBA code is written.

aiming-lab/AutoResearchClaw15k—~1.9kAutomated safety check: PassMIT1 mo ago
9

Plans, runs, and documents analytical method validation, verification, or transfer studies under ICH Q2(R2)/Q14, USP, ICH M10, CLSI EP, or ISO/IEC 17025.

K-Dense-AI/scientific-agent-skills48k1 repo~4.9kAutomated safety check: NotesMIT2 days ago
10

A skill your agent uses when the user asks to "design an A/B test", "set up a creative/landing test", "run an incrementality test", or "is this result statistically and practically material?"…

aaron-he-zhu/aaron-marketing-skills2.9k2 repos~2.8kAutomated safety check: PassApache-2.0today
11

Critically review strategy drafts from edge-strategy-designer for edge plausibility, overfitting risk, sample size adequacy, and execution realism.

tradermonty/claude-trading-skills3k1 repo~988Automated safety check: PassMITyesterday
12

Turns a vague research direction into a focused, problem-anchored method plan through up to five review rounds with a second model.

zjYao36/Auto-Research-Refine1287 repos~6.9kAutomated safety check: NotesNo licence6 mo ago
13

World-class data science skill for statistical modeling, experimentation, causal inference, and advanced analytics.

Raidriar7170/hermes-skilleval1256 repos~1.4kAutomated safety check: PassMIT11 days ago
14

Design experiments and studies BEFORE data is collected — choosing a design, randomizing, blocking, and laying out treatment combinations so results are interpretable.

Oleafly/Oleafly2053 repos~3.5kAutomated safety check: NotesMITtoday
15

Experiment executor and monitor for academic research. An agent skill from Imbad0202/experiment-agent.

Imbad0202/experiment-agent199—~3.1kAutomated safety check: PassCC-BY-NC-4.012 days ago
16

Statistical significance calculator for A/B test results with sample size requirements, segment breakdowns, and hypothesis generation.

irinabuht12-oss/marketing-skills3.8k—~1.4kAutomated safety check: PassNo licence13 days ago
17

Systematic framework for evaluating scholarly and research work based on the ScholarEval methodology.

jimmc414/Kosmos5942 repos~2.5kAutomated safety check: PassNo licence3 days ago
18

Facilitates evidence-aware scientific ideation with independent generation, structured discussion, explicit assumptions, transparent evaluation, adversarial review, and decision logs.

Oleafly/Oleafly2052 repos~3.5kAutomated safety check: PassMITtoday
19

Defines a testable hypothesis with clear success metrics and a validation approach.

product-on-purpose/pm-skills713—~966Automated safety check: PassApache-2.02 days ago
20

Convert a completed data-analysis conversation into evidence-backed, reproducible living research through the Krisk MCP server.

napjon/krisk118—~702Automated safety check: PassBSD-3-Clause5 days ago
21

Design an A/B experiment — hypothesis, variants, primary metric, and sample size.

Owl-Listener/designer-skills2.9k1 repo~472Automated safety check: PassMIT1 mo ago
22

ML paper pipeline: experiment design to submission. An agent skill from HezaoHezao/poirot.

HezaoHezao/poirot2501 repo~1.4kAutomated safety check: PassMIT2 mo ago
23

Design a statistically rigorous A/B or multivariate test plan — If/Then/Because hypothesis, control and variant specs, required sample size per variant (absolute vs relative MDE via…

indranilbanerjee/digital-marketing-pro8541 repo~2kAutomated safety check: PassMIT3 days ago
24

Complete academic research skill suite covering the full pipeline: paper reading (read/explain papers with storytelling), idea generation (brainstorm research directions), experiment design (plan…

voidful/academic-skills132—~887Automated safety check: PassMIT6 mo ago
25

Best practices for systematic research, source evaluation, and evidence gathering

vstorm-co/pydantic-deepagents1.1k—~577Automated safety check: PassMITyesterday
26

Produces ranked, evidence-grounded next steps when an ML experiment has stalled, drawing on a Leeroopedia knowledge base or on fetched docs and issues.

Leeroo-AI/superml195—~4.8kAutomated safety check: PassApache-2.06 mo ago
27

Plans research experiments in four progressive stages, from a first working implementation through baseline tuning and creative research to ablation studies.

lingzhi227/agent-research-skills383—~752Automated safety check: PassNo licence7 mo ago
28

PhD-level expertise in data science, statistics, and machine learning.

magnus919/hermes-profiles278—~3.3kAutomated safety check: PassMIT3 mo ago
29

Methodology for market research and data collection, ensuring data quality and source traceability

chekusu/wanman688—~533Automated safety check: PassApache-2.03 mo ago
30

A skill your agent uses when the user wants to design experiments, plan ablation studies, structure baselines, or create incremental evaluation strategies.

fcakyon/phd-skills414—~987Automated safety check: PassMIT21 days ago
31

學術研究實驗設計技能——從研究假設到可重現實驗計畫的完整流程。當使用者需要規劃實驗、設計 ablation study、選擇 baseline、確定評估指標,或問「我應該跑哪些實驗」時,一定要使用此技能。觸發詞包括:實驗設計、experiment design、ablation、baseline、跑什麼實驗、evaluation metric、如何驗證方法。適用於機器學習、NLP、CV…

voidful/academic-skills132—~1.2kAutomated safety check: PassMIT6 mo ago
32

Stress-test an academic research question, proposal, study design, analysis plan, manuscript claim, review protocol, or AI research project through a one-question-at-a-time interview until its…

Exekiel179/psyclaw103—~2kAutomated safety check: PassMIT9 days ago
33

Search scientific papers and retrieve structured experimental data extracted from full-text studies via the BGPT MCP server.

majiayu000/claude-skill-registry6664 repos~619Automated safety check: NotesMITtoday
34

Applies the epidemiological reasoning and population-health frameworks of Albert Hofman (Harvard epidemiologist, Rotterdam Study).

K-Dense-AI/mimeographs129—~1.4kAutomated safety check: PassMIT1 mo ago
35

[production-grade internal] Optimizes AI/ML/LLM usage when you need model selection, prompt engineering, cost reduction, or experiment design.

nagisanzenin/production-grade181—~3kAutomated safety check: PassNo licence1 mo ago
36

A skill your agent uses when checking a radiology or medical AI study design before drafting or submission.

Aperivue/medsci-skills329—~3.9kAutomated safety check: PassMIT2 days ago
37

A skill your agent uses when designing or analyzing a controlled experiment — falsifiable hypothesis, sample size from an MDE, reading significance/CI/power, CUPED, or rescuing tests that won't go…

ericrisco/rsc-harness156—~2.4kAutomated safety check: PassMITtoday
38

Multiagent AI system for scientific research assistance that automates research workflows from data analysis to publication.

davila7/claude-code-templates32k9 repos~1.5kAutomated safety check: NotesMITtoday
39

Calculate A/B test statistical significance. An agent skill from guia-matthieu/clawfu-skills.

guia-matthieu/clawfu-skills150—~1kAutomated safety check: PassMIT6 days ago
40

A skill your agent uses when the user needs help with Vivado simulation strategy, flow selection, and debugging.

Shinei-Nouzen-Arch/FPGA-Agent178—~2.9kAutomated safety check: PassGPL-2.01 mo ago
41

Mine and synthesize real top-journal scenario/vignette experiment patterns for behavioral research.

Drchronx/ai-agent-research-starter-kit134—~1.1kAutomated safety check: PassUnknown4 mo ago
42

Design, plan, and analyze A/B tests with statistical rigor. An agent skill from OpenClaudia/openclaudia-skills.

OpenClaudia/openclaudia-skills708—~1.7kAutomated safety check: PassMIT19 days ago
43

Generate falsifiable trade strategy hypotheses from market data, trade logs, and journal snippets.

tradermonty/claude-trading-skills3k1 repo~693Automated safety check: PassMITyesterday
44

Design/audit tissue, cell, organoid, animal, perturbation and rescue validation; not physical SOP execution.

huang-sir1/radiology-skills1.9k—~3.6kAutomated safety check: PassUnknown16 days ago
45

Guides turning an observation into testable, falsifiable hypotheses with null and alternative statements, competing explanations and predictions tied to experimental design.

aiming-lab/AutoResearchClaw15k—~628Automated safety check: PassMIT1 mo ago
46

AlphaEvolve expert consultant grounded strictly in the official reference guide.

Google-Cloud-AI/alphaevolve-on-googlecloud118—~3kAutomated safety check: PassApache-2.06 days ago
47

Causal inference toolkit for when experiments are not possible: estimate treatment effects from observational data with assumption checks and mandatory caveats.

ai-analyst-lab/ai-analyst304—~1.8kAutomated safety check: PassMIT6 days ago
48

Frames an epic as a testable if-then hypothesis with a target user, expected outcome, small discovery experiments and validation measures.

deanpeters/Product-Manager-Skills7.2k1 repo~3.2kAutomated safety check: PassUnknown1 mo ago

Questions, answered from the data.

What is the best experimental design skill?

Scientific Critical Thinking from weapp-tailwindcss/weapp-tailwindcss ranks first of the 286 experimental design skills listed here, with the highest score: its repository has 1.9k GitHub stars, 23 other GitHub owners carry a copy, its SKILL.md loads about 5.9k tokens and it has informational notes only in the automated safety check. Next come Statistical Analysis and Claim-Driven Experiment Planner.

Which experimental design skills are official?

2 of the 286 experimental design skills are official, published by the vendor's own GitHub organization: Clinical Protocol Drafting and Content Experimentation Best Practices.

How are these skills ranked?

By Skill Navigator score, which combines the GitHub stars of the skill's repository (shared across that repo's skills and discounted for large collections), how many other GitHub owners carry a copy of the skill, and automated SKILL.md quality checks, minus penalties for safety-check warnings and for each further skill from the same repository. Skills that fail the safety check are not listed.