Topic · Research & Science
Best experimental design skills for Claude Code, Codex and other agents.
- skills
- 286
- official
- 2
Experimental design skills, ranked
Ranked by score. Sort bymost stars,trending,newest,recently updated
| # | Skill | Repository | Stars | Used in | Tokens | Auto-check | Licence | Updated |
|---|---|---|---|---|---|---|---|---|
| 1 | Evaluate research rigor. An agent skill from weapp-tailwindcss/weapp-tailwindcss. | weapp-tailwindcss/ | 1.9k | 23 repos | ~5.9k | Automated safety check: Notes | MIT | today |
| 2 | Guided statistical analysis for research data - test selection, assumption checking, effect sizes, power analysis, Bayesian alternatives, and APA-formatted reporting. | spacering-net/ | 3.8k | 3 repos | ~5k | Automated safety check: Pass | MIT | today |
| 3 | Turns a refined research proposal into a claim-to-evidence-to-run-order roadmap instead of a sprawling benchmark wishlist. | zjYao36/ | 128 | 7 repos | ~2.3k | Automated safety check: Notes | No licence | 6 mo ago |
| 4 | Structures benchmark and evaluation papers around five pillars, with a completeness audit, an Introduction logic chain, a section skeleton and a pre-submission checklist. | HKUSTDial/ | 8.4k | — | ~2.8k | Automated safety check: Pass | CC-BY-4.0 | 1 mo ago |
| 5 | Sample-size and statistical power calculations for planning studies. | spacering-net/ | 3.8k | 1 repo | ~3.6k | Automated safety check: Notes | MIT | today |
| 6 | Chains research-refine and experiment-plan to turn a vague research direction into a focused proposal and a claim-driven experiment roadmap. | zjYao36/ | 128 | 6 repos | ~1.4k | Automated safety check: Notes | No licence | 6 mo ago |
| 7 | Builds a scoring-aligned outline for a mathematical modeling paper and a model selection plan with baseline, improvement and validation experiments. | yushui2022/ | 452 | 1 repo | ~1.8k | Automated safety check: Pass | MIT | 10 days ago |
| 8 | Turns a broad metabolic modelling topic into a concrete, paper-shaped plan with organism, model, perturbations, metrics and figures before any FBA code is written. | aiming-lab/ | 15k | — | ~1.9k | Automated safety check: Pass | MIT | 1 mo ago |
| 9 | Plans, runs, and documents analytical method validation, verification, or transfer studies under ICH Q2(R2)/Q14, USP, ICH M10, CLSI EP, or ISO/IEC 17025. | K-Dense-AI/ | 48k | 1 repo | ~4.9k | Automated safety check: Notes | MIT | 2 days ago |
| 10 | A skill your agent uses when the user asks to "design an A/B test", "set up a creative/landing test", "run an incrementality test", or "is this result statistically and practically material?"… | aaron-he-zhu/ | 2.9k | 2 repos | ~2.8k | Automated safety check: Pass | Apache-2.0 | today |
| 11 | Critically review strategy drafts from edge-strategy-designer for edge plausibility, overfitting risk, sample size adequacy, and execution realism. | tradermonty/ | 3k | 1 repo | ~988 | Automated safety check: Pass | MIT | yesterday |
| 12 | Turns a vague research direction into a focused, problem-anchored method plan through up to five review rounds with a second model. | zjYao36/ | 128 | 7 repos | ~6.9k | Automated safety check: Notes | No licence | 6 mo ago |
| 13 | World-class data science skill for statistical modeling, experimentation, causal inference, and advanced analytics. | Raidriar7170/ | 125 | 6 repos | ~1.4k | Automated safety check: Pass | MIT | 11 days ago |
| 14 | Design experiments and studies BEFORE data is collected — choosing a design, randomizing, blocking, and laying out treatment combinations so results are interpretable. | Oleafly/ | 205 | 3 repos | ~3.5k | Automated safety check: Notes | MIT | today |
| 15 | Experiment executor and monitor for academic research. An agent skill from Imbad0202/experiment-agent. | Imbad0202/ | 199 | — | ~3.1k | Automated safety check: Pass | CC-BY-NC-4.0 | 12 days ago |
| 16 | Statistical significance calculator for A/B test results with sample size requirements, segment breakdowns, and hypothesis generation. | irinabuht12-oss/ | 3.8k | — | ~1.4k | Automated safety check: Pass | No licence | 13 days ago |
| 17 | Systematic framework for evaluating scholarly and research work based on the ScholarEval methodology. | jimmc414/ | 594 | 2 repos | ~2.5k | Automated safety check: Pass | No licence | 3 days ago |
| 18 | Facilitates evidence-aware scientific ideation with independent generation, structured discussion, explicit assumptions, transparent evaluation, adversarial review, and decision logs. | Oleafly/ | 205 | 2 repos | ~3.5k | Automated safety check: Pass | MIT | today |
| 19 | Defines a testable hypothesis with clear success metrics and a validation approach. | product-on-purpose/ | 713 | — | ~966 | Automated safety check: Pass | Apache-2.0 | 2 days ago |
| 20 | Convert a completed data-analysis conversation into evidence-backed, reproducible living research through the Krisk MCP server. | napjon/ | 118 | — | ~702 | Automated safety check: Pass | BSD-3-Clause | 5 days ago |
| 21 | Design an A/B experiment — hypothesis, variants, primary metric, and sample size. | Owl-Listener/ | 2.9k | 1 repo | ~472 | Automated safety check: Pass | MIT | 1 mo ago |
| 22 | ML paper pipeline: experiment design to submission. An agent skill from HezaoHezao/poirot. | HezaoHezao/ | 250 | 1 repo | ~1.4k | Automated safety check: Pass | MIT | 2 mo ago |
| 23 | 23.Ab Test Plan Design a statistically rigorous A/B or multivariate test plan — If/Then/Because hypothesis, control and variant specs, required sample size per variant (absolute vs relative MDE via… | indranilbanerjee/ | 854 | 1 repo | ~2k | Automated safety check: Pass | MIT | 3 days ago |
| 24 | Complete academic research skill suite covering the full pipeline: paper reading (read/explain papers with storytelling), idea generation (brainstorm research directions), experiment design (plan… | voidful/ | 132 | — | ~887 | Automated safety check: Pass | MIT | 6 mo ago |
| 25 | Best practices for systematic research, source evaluation, and evidence gathering | vstorm-co/ | 1.1k | — | ~577 | Automated safety check: Pass | MIT | yesterday |
| 26 | Produces ranked, evidence-grounded next steps when an ML experiment has stalled, drawing on a Leeroopedia knowledge base or on fetched docs and issues. | Leeroo-AI/ | 195 | — | ~4.8k | Automated safety check: Pass | Apache-2.0 | 6 mo ago |
| 27 | Plans research experiments in four progressive stages, from a first working implementation through baseline tuning and creative research to ablation studies. | lingzhi227/ | 383 | — | ~752 | Automated safety check: Pass | No licence | 7 mo ago |
| 28 | PhD-level expertise in data science, statistics, and machine learning. | magnus919/ | 278 | — | ~3.3k | Automated safety check: Pass | MIT | 3 mo ago |
| 29 | Methodology for market research and data collection, ensuring data quality and source traceability | chekusu/ | 688 | — | ~533 | Automated safety check: Pass | Apache-2.0 | 3 mo ago |
| 30 | A skill your agent uses when the user wants to design experiments, plan ablation studies, structure baselines, or create incremental evaluation strategies. | fcakyon/ | 414 | — | ~987 | Automated safety check: Pass | MIT | 21 days ago |
| 31 | 學術研究實驗設計技能——從研究假設到可重現實驗計畫的完整流程。當使用者需要規劃實驗、設計 ablation study、選擇 baseline、確定評估指標,或問「我應該跑哪些實驗」時,一定要使用此技能。觸發詞包括:實驗設計、experiment design、ablation、baseline、跑什麼實驗、evaluation metric、如何驗證方法。適用於機器學習、NLP、CV… | voidful/ | 132 | — | ~1.2k | Automated safety check: Pass | MIT | 6 mo ago |
| 32 | Stress-test an academic research question, proposal, study design, analysis plan, manuscript claim, review protocol, or AI research project through a one-question-at-a-time interview until its… | Exekiel179/ | 103 | — | ~2k | Automated safety check: Pass | MIT | 9 days ago |
| 33 | Search scientific papers and retrieve structured experimental data extracted from full-text studies via the BGPT MCP server. | majiayu000/ | 666 | 4 repos | ~619 | Automated safety check: Notes | MIT | today |
| 34 | Applies the epidemiological reasoning and population-health frameworks of Albert Hofman (Harvard epidemiologist, Rotterdam Study). | K-Dense-AI/ | 129 | — | ~1.4k | Automated safety check: Pass | MIT | 1 mo ago |
| 35 | [production-grade internal] Optimizes AI/ML/LLM usage when you need model selection, prompt engineering, cost reduction, or experiment design. | nagisanzenin/ | 181 | — | ~3k | Automated safety check: Pass | No licence | 1 mo ago |
| 36 | 36.Design Study A skill your agent uses when checking a radiology or medical AI study design before drafting or submission. | Aperivue/ | 329 | — | ~3.9k | Automated safety check: Pass | MIT | 2 days ago |
| 37 | 37.Ab Testing A skill your agent uses when designing or analyzing a controlled experiment — falsifiable hypothesis, sample size from an MDE, reading significance/CI/power, CUPED, or rescuing tests that won't go… | ericrisco/ | 156 | — | ~2.4k | Automated safety check: Pass | MIT | today |
| 38 | 38.Denario Multiagent AI system for scientific research assistance that automates research workflows from data analysis to publication. | davila7/ | 32k | 9 repos | ~1.5k | Automated safety check: Notes | MIT | today |
| 39 | Calculate A/B test statistical significance. An agent skill from guia-matthieu/clawfu-skills. | guia-matthieu/ | 150 | — | ~1k | Automated safety check: Pass | MIT | 6 days ago |
| 40 | 40.Vivado Sim A skill your agent uses when the user needs help with Vivado simulation strategy, flow selection, and debugging. | Shinei-Nouzen-Arch/ | 178 | — | ~2.9k | Automated safety check: Pass | GPL-2.0 | 1 mo ago |
| 41 | Mine and synthesize real top-journal scenario/vignette experiment patterns for behavioral research. | Drchronx/ | 134 | — | ~1.1k | Automated safety check: Pass | Unknown | 4 mo ago |
| 42 | Design, plan, and analyze A/B tests with statistical rigor. An agent skill from OpenClaudia/openclaudia-skills. | OpenClaudia/ | 708 | — | ~1.7k | Automated safety check: Pass | MIT | 19 days ago |
| 43 | Generate falsifiable trade strategy hypotheses from market data, trade logs, and journal snippets. | tradermonty/ | 3k | 1 repo | ~693 | Automated safety check: Pass | MIT | yesterday |
| 44 | Design/audit tissue, cell, organoid, animal, perturbation and rescue validation; not physical SOP execution. | huang-sir1/ | 1.9k | — | ~3.6k | Automated safety check: Pass | Unknown | 16 days ago |
| 45 | Guides turning an observation into testable, falsifiable hypotheses with null and alternative statements, competing explanations and predictions tied to experimental design. | aiming-lab/ | 15k | — | ~628 | Automated safety check: Pass | MIT | 1 mo ago |
| 46 | AlphaEvolve expert consultant grounded strictly in the official reference guide. | Google-Cloud-AI/ | 118 | — | ~3k | Automated safety check: Pass | Apache-2.0 | 6 days ago |
| 47 | 47.Causal Causal inference toolkit for when experiments are not possible: estimate treatment effects from observational data with assumption checks and mandatory caveats. | ai-analyst-lab/ | 304 | — | ~1.8k | Automated safety check: Pass | MIT | 6 days ago |
| 48 | Frames an epic as a testable if-then hypothesis with a target user, expected outcome, small discovery experiments and validation measures. | deanpeters/ | 7.2k | 1 repo | ~3.2k | Automated safety check: Pass | Unknown | 1 mo ago |
Questions, answered from the data.
What is the best experimental design skill?
Scientific Critical Thinking from weapp-tailwindcss/weapp-tailwindcss ranks first of the 286 experimental design skills listed here, with the highest score: its repository has 1.9k GitHub stars, 23 other GitHub owners carry a copy, its SKILL.md loads about 5.9k tokens and it has informational notes only in the automated safety check. Next come Statistical Analysis and Claim-Driven Experiment Planner.
Which experimental design skills are official?
2 of the 286 experimental design skills are official, published by the vendor's own GitHub organization: Clinical Protocol Drafting and Content Experimentation Best Practices.
How are these skills ranked?
By Skill Navigator score, which combines the GitHub stars of the skill's repository (shared across that repo's skills and discounted for large collections), how many other GitHub owners carry a copy of the skill, and automated SKILL.md quality checks, minus penalties for safety-check warnings and for each further skill from the same repository. Skills that fail the safety check are not listed.
Explore related skills
Category
More topics in Research & Science
- Bioinformatics1,169
- Citation management952
- Academic paper search515
- Literature review487
- Deep research421
- Reproducible research378
- Econometrics and empirical research331
- Peer review314
- Clinical and healthcare research271
- Scientific writing258
- Hypothesis generation224
- Physical and earth sciences213
- Drug discovery and cheminformatics202
- Fact-checking and source verification174
- Protein structure and design132
- Math and symbolic computation64
- Grant writing55
- Quantum computing24