Agent skill

Research Experiment Designer

by lingzhi227 in lingzhi227/agent-research-skills

Plans research experiments in four progressive stages, from a first working implementation through baseline tuning and creative research to ablation studies.

No licenceAuto-check passedResearch & Science

Install Research Experiment Designer

skills CLI
$ npx skills add lingzhi227/agent-research-skills --skill experiment-design -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install lingzhi227/agent-research-skills experiment-design --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/lingzhi227/agent-research-skills.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/experiment-design .claude/skills/experiment-design && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
experiment-design
GitHub stars
390
Token cost
~752 tokens
SKILL.md length
200 words
Files
3 (incl. scripts, references)
Skills in repo
31
Repo updated
First seen
Licence
None found

At a glance

Plans research experiments in four progressive stages, from a first working implementation through baseline tuning and creative research to ablation studies.

  • Works in 4 steps: Initial Implementation → Baseline Tuning → Creative Research → …
  • Planning experiments for a research paper
  • SKILL.md covers Input, References, Scripts and 4-Stage Progressive Framework…, plus 3 more sections
  • Runs Python scripts from its folder; calls python

What it does

Given a research idea, plan or method description, the skill structures an experiment plan around four stages taken from AI-Scientist-v2. Stage 1 gets a simple working implementation on a basic dataset, stage 2 tunes hyperparameters without changing the architecture and tests on at least two datasets, stage 3 explores novel improvements on at least three, and stage 4 runs systematic ablations on the same datasets.

A bundled stdlib-only Python script, `scripts/design_experiments.py`, generates baselines, an ablation matrix, a hyperparameter grid and metric selection from a plan JSON file or from a method and task, with JSON or Markdown output. The rules call for multi-seed evaluation, logging every run in notes.txt and figures for training curves. A references file holds the stage prompts.

When your agent uses it

  • Planning experiments for a research paper
  • Choosing baselines and an ablation matrix for a new method
  • Laying out hyperparameter sweeps and evaluation metrics

Example prompts

  • “Design an experiment plan for contrastive learning on image classification.”
  • “Generate an ablation matrix and hyperparameter grid from research_plan.json.”
  • “Which baselines and metrics should I use for my graph-based recommender idea?”

Requirements

  • Python, standard library only, to run `scripts/design_experiments.py`

Workflow steps

4 steps, taken from the step headings in SKILL.md.

  1. Initial Implementation
  2. Baseline Tuning
  3. Creative Research
  4. Ablation Studies

What it can do on your machine

Read from SKILL.md and the folder at commit 9e6c085. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Ships 1 file in scripts/ (Python), which the agent can run.

    Shell commands in SKILL.md call:

    • python

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Research Experiment Designer loads about 752 tokens when it runs, and up to ~1.6k if it reads all its reference files. Until then it costs about 69 tokens; SKILL.md has 200 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~69
When it runs · the whole SKILL.md, loaded when a task matches
~752
With references · SKILL.md plus every file in references/, read only if the agent opens them
~1.6k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); the scripts in this folder are not scanned.

SKILL.md

Without a licence we can't republish the file, so here is its outline and opening line. It has 200 words (~752 tokens).

“Design structured, progressive experiment plans for research papers.”

— opening of SKILL.md by lingzhi227
name
experiment-design
argument-hint
idea-or-plan

Read the full SKILL.md on GitHub

Files

SKILL.md and 2 other files (scripts, references) in skills/experiment-design of lingzhi227/agent-research-skills.

  • SKILL.md
  • references/stage-prompts.md
  • scripts/design_experiments.py

Open the folder on GitHubat commit 9e6c085

Compare with similar skills

Research Experiment Designer next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Research Experiment Designer compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Research Experiment Designer this skilllingzhi227/agent-research-skills390—~752Automated safety check: PassNone
Scientific Critical Thinkingweapp-tailwindcss/weapp-tailwindcss1.9k22 repos~5.9kAutomated safety check: NotesMIT
Benchmark Paper TemplateHKUSTDial/Supervisor-Skills8.8k—~2.8kAutomated safety check: PassCC-BY-4.0
Claim-Driven Experiment PlannerzjYao36/Auto-Research-Refine1286 repos~2.3kAutomated safety check: NotesNone
Research Refine PipelinezjYao36/Auto-Research-Refine1285 repos~1.4kAutomated safety check: NotesNone
Metabolic Study Planneraiming-lab/AutoResearchClaw15k—~1.9kAutomated safety check: PassMIT

Similar skills

  • Scientific Critical Thinking

    weapp-tailwindcss/weapp-tailwindcss

    Evaluate research rigor. An agent skill from weapp-tailwindcss/weapp-tailwindcss.

    1.9k GitHub starsUsed in 22 repos~5.9k tokens
    Research & ScienceAuto-check: notes
  • Benchmark Paper Template

    HKUSTDial/Supervisor-Skills

    Structures benchmark and evaluation papers around five pillars, with a completeness audit, an Introduction logic chain, a section skeleton and a pre-submission checklist.

    8.8k GitHub stars~2.8k tokensUpdated 1 mo ago
    Research & ScienceAuto-check passed
  • Claim-Driven Experiment Planner

    zjYao36/Auto-Research-Refine

    Turns a refined research proposal into a claim-to-evidence-to-run-order roadmap instead of a sprawling benchmark wishlist.

    128 GitHub starsUsed in 6 repos~2.3k tokens
    Research & ScienceAuto-check: notes
  • Research Refine Pipeline

    zjYao36/Auto-Research-Refine

    Chains research-refine and experiment-plan to turn a vague research direction into a focused proposal and a claim-driven experiment roadmap.

    128 GitHub starsUsed in 5 repos~1.4k tokens
    Research & ScienceAuto-check: notes
  • Metabolic Study Planner

    aiming-lab/AutoResearchClaw

    Turns a broad metabolic modelling topic into a concrete, paper-shaped plan with organism, model, perturbations, metrics and figures before any FBA code is written.

    15k GitHub stars~1.9k tokensUpdated 1 mo ago
    Research & ScienceAuto-check passed
  • Analytical Method Validation Planner

    K-Dense-AI/scientific-agent-skills

    Plans, runs, and documents analytical method validation, verification, or transfer studies under ICH Q2(R2)/Q14, USP, ICH M10, CLSI EP, or ISO/IEC 17025.

    48k GitHub starsUsed in 1 repo~4.9k tokens
    Research & ScienceAuto-check: notes

More from lingzhi227/agent-research-skills

All 31 skills in this repo
  • Backward Traceability

    lingzhi227/agent-research-skills

    Makes each number in a LaTeX paper link back to the code line that produced it, using hypertarget and hyperlink tags and compile-time `\num` formulas.

    390 GitHub stars~802 tokensUpdated 7 mo ago
    Auto-check passed
  • Statistical Data Analysis

    lingzhi227/agent-research-skills

    Writes statistical analysis code for experimental data, runs it through a four-round review, and reports effect sizes, p-values and confidence intervals.

    390 GitHub stars~886 tokensUpdated 7 mo ago
    Auto-check passed
  • Excalidraw Canvas Toolkit

    lingzhi227/agent-research-skills

    Draws and refines Excalidraw diagrams on a live canvas through MCP tools or a REST API, with screenshots, file import and export, snapshots and Mermaid conversion.

    390 GitHub stars~3.8k tokensUpdated 7 mo ago
    Auto-check passed
  • Scientific Figure Generation

    lingzhi227/agent-research-skills

    Generates publication-quality scientific figures with matplotlib or seaborn through query expansion, a run-and-retry coding loop and a visual check of the rendered PNG.

    390 GitHub stars~809 tokensUpdated 7 mo ago
    Auto-check passed
  • Research Idea Generation

    lingzhi227/agent-research-skills

    Generates and iteratively refines research ideas for a given area, checking each one's novelty against Semantic Scholar and arXiv, and scoring it on interestingness, feasibility and novelty.

    390 GitHub stars~747 tokensUpdated 7 mo ago
    Auto-check passed
  • Academic LaTeX Formatter

    lingzhi227/agent-research-skills

    Sets up conference-specific LaTeX paper templates, checks a draft for formatting and submission issues, and auto-fixes common problems for venues like ICML, ICLR, NeurIPS, AAAI and ACL.

    390 GitHub stars~603 tokensUpdated 7 mo ago
    Auto-check passed

Questions about Research Experiment Designer

What does Research Experiment Designer do?

Plans research experiments in four progressive stages, from a first working implementation through baseline tuning and creative research to ablation studies. Given a research idea, plan or method description, the skill structures an experiment plan around four stages taken from AI-Scientist-v2. Stage 1 gets a simple working implementation on a basic dataset, stage 2 tunes hyperparameters without changing the architecture and tests on at least two datasets, stage 3 explores novel improvements on at least three, and stage 4 runs systematic ablations on the same datasets.

When should I use Research Experiment Designer?

Research Experiment Designer fits situations like: planning experiments for a research paper; choosing baselines and an ablation matrix for a new method; laying out hyperparameter sweeps and evaluation metrics.

How do I install Research Experiment Designer in Claude Code?

Run `npx skills add lingzhi227/agent-research-skills --skill experiment-design -a claude-code`. Or copy the skill folder (skills/experiment-design in lingzhi227/agent-research-skills) into .claude/skills/experiment-design in your project. Claude Code loads it when a task matches its description.

How do I install Research Experiment Designer in Codex?

Run `npx skills add lingzhi227/agent-research-skills --skill experiment-design -a codex`. Or copy the skill folder (skills/experiment-design in lingzhi227/agent-research-skills) into .agents/skills/experiment-design in your project. Codex loads it when a task matches its description.

Can I use Research Experiment Designer in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add lingzhi227/agent-research-skills --skill experiment-design -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/experiment-design, .gemini/skills/experiment-design, .github/skills/experiment-design and .opencode/skills/experiment-design in your project.

What does Research Experiment Designer need to run?

Going by SKILL.md and its folder, Research Experiment Designer needs Python for the scripts in its folder and the command-line tools its instructions call (python). Our summary lists: Python, standard library only, to run `scripts/design_experiments.py`.

Does Research Experiment Designer access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is Research Experiment Designer safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. The check reads SKILL.md only: the scripts in the folder are not scanned, so read them before running anything.

What licence does Research Experiment Designer use?

No licence was found for Research Experiment Designer or its repository. Without one, default copyright applies: ask the author before reusing or redistributing it.

How many tokens does Research Experiment Designer use?

About 752 tokens (SKILL.md is roughly 3k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full. Its references folder adds about 878 tokens, read only when the agent opens those files.

What are the alternatives to Research Experiment Designer?

Skills that share tags, products or a category with Research Experiment Designer: Scientific Critical Thinking (weapp-tailwindcss/weapp-tailwindcss, 1.9k stars), Benchmark Paper Template (HKUSTDial/Supervisor-Skills, 8.8k stars), Claim-Driven Experiment Planner (zjYao36/Auto-Research-Refine, 128 stars) and Research Refine Pipeline (zjYao36/Auto-Research-Refine, 128 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Research Experiment Designer?

lingzhi227 (a GitHub user) maintains it in lingzhi227/agent-research-skills, which has 390 GitHub stars. The repository holds 31 skills in this directory. The repository was last updated on February 27, 2026.

Source: lingzhi227/agent-research-skills on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.