Agent skill

Run Experiment

by AI4Scientist in AI4Scientist/nano-scientist

Deploy and run ML experiments on local, remote, Vast.ai, or Modal serverless GPU.

No licenceAuto-check: notesBackend & APIs

Install Run Experiment

skills CLI
$ npx skills add AI4Scientist/nano-scientist --skill run-experiment -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install AI4Scientist/nano-scientist run-experiment --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/AI4Scientist/nano-scientist.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/run-experiment .claude/skills/run-experiment && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
run-experiment
GitHub stars
128
Used in
4 other repos
Token cost
~2.9k tokens
SKILL.md length
977 words
Files
1
Skills in repo
74
Repo updated
First seen
Licence
None found

At a glance

Deploy and run ML experiments on local, remote, Vast.ai, or Modal serverless GPU.

  • Works in 8 steps: Detect Environment → Pre-flight Check → Sync Code (Remote Only) → …
  • User says run experiment
  • SKILL.md covers Workflow, Key Rules and CLAUDE.md Example
  • Calls ssh, python and git; reaches wandb.ai; needs WANDB_API_KEY

What it does

Run Experiment is an agent skill from AI4Scientist/nano-scientist. Deploy and run ML experiments on local, remote, Vast.ai, or Modal serverless GPU. Use when user says "run experiment", "deploy to server", "跑实验", or needs to launch training jobs.

Its SKILL.md is about 2.9k tokens, which your agent loads only when the skill is triggered. It is a single SKILL.md file with no bundled scripts.

It sits in Backend & APIs, covering Serverless. The repository describes itself as: An autonomous research agent that turns a topic into a peer-reviewed technical report.

When your agent uses it

  • User says run experiment
  • Deploy to server
  • Needs to launch training jobs

Example prompts

  • “run experiment”
  • “deploy to server”
  • “/run-experiment”

Requirements

  • Python 3
  • Docker
  • A credential in WANDB_API_KEY
  • A credential in YOUR_KEY
  • Pre-approved tools (allowed-tools): Bash(*), Read, Grep, Glob, Edit, Write, Agent, Skill(serverless-modal)

Workflow steps

8 steps, taken from the step headings in SKILL.md.

  1. Detect Environment
  2. Pre-flight Check
  3. Sync Code (Remote Only)
  4. 5: W&B Integration (when wandb: true in CLAUDE.md)
  5. Deploy
  6. Verify Launch
  7. Feishu Notification (if configured)
  8. Auto-Destroy Vast.ai Instance (when gpu: vast and auto_destroy: true)

What it can do on your machine

Read from SKILL.md and the folder at commit 7132192. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves these tools, so the agent can use them without asking each time:

    • Bash(*)
    • Read
    • Grep
    • Glob
    • Edit
    • Write
    • Agent
    • Skill(serverless-modal)

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Shell commands in SKILL.md call:

    • ssh
    • python
    • git
    • modal
    • rsync
    • scp
    • conda
    • pip

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    Hosts in commands or code, which the agent is likely to contact:

    • wandb.ai

    Also links to:

    • cloud.vast.ai
    • modal.com

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names these keys or tokens, usually read from environment variables:

    • WANDB_API_KEY

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Run Experiment loads about 2.9k tokens when it runs. Until then it costs about 49 tokens; SKILL.md has 977 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~49
When it runs · the whole SKILL.md, loaded when a task matches
~2.9k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check: notes

The automated check noted patterns worth knowing about, such as sudo or a known installer.

  • NotePre-approves every shell command (allowed-tools: Bash)SKILL.md
    allowed-tools: Bash(*), Read, Grep, Glob, Edit, Write, Agent, Skill(serverless-modal)

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

Without a licence we can't republish the file, so here is its outline and opening line. It has 977 words (~2,946 tokens).

“Deploy and run ML experiment: $ARGUMENTS”

— opening of SKILL.md by AI4Scientist
name
run-experiment
allowed-tools
Bash(*), Read, Grep, Glob, Edit, Write, Agent, Skill(serverless-modal)
argument-hint
experiment-description

Read the full SKILL.md on GitHub

Files

Just SKILL.md in skills/run-experiment of AI4Scientist/nano-scientist.

Open the folder on GitHubat commit 7132192

Used in 4 other repositories

We found 9 copies of this SKILL.md (exact, near-identical or edited) in other folders, from 4 other GitHub owners. This page covers the copy in AI4Scientist/nano-scientist, which our catalogue first saw on October 7, 2026.

Compare with similar skills

Run Experiment next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Run Experiment compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Run Experiment this skillAI4Scientist/nano-scientist1284 repos~2.9kAutomated safety check: NotesNone
Arcgis To Portaljsdatopian/portaljs2.4k1 repos~2kAutomated safety check: PassMIT
AWS Serverless Edazxkane/aws-skills3674 repos~3.2kAutomated safety check: PassMIT
NubaseOtterMind/Nubase623—~2.2kAutomated safety check: NotesApache-2.0
Tianji Worker Operationsmsgbyte/tianji3.1k—~1.1kAutomated safety check: PassApache-2.0
AI Model NodejsTencentCloudBase/CloudBase-AI-Toolkit1.1k2 repos~5kAutomated safety check: PassMIT

Similar skills

  • Arcgis To Portaljs

    datopian/portaljs

    Migrate a whole ArcGIS Hub site into a PortalJS Arc portal end-to-end.

    2.4k GitHub starsUsed in 1 repo~2k tokens
    Backend & APIsAuto-check passed
  • AWS Serverless Eda

    zxkane/aws-skills

    AWS serverless and event-driven architecture expert based on Well-Architected Framework.

    367 GitHub starsUsed in 4 repos~3.2k tokens
    Backend & APIsAuto-check passed
  • Nubase

    OtterMind/Nubase

    A skill your agent uses when the user mentions Nubase broadly, wants a backend for an AI-generated app, or needs to deploy/publish generated code online — across Database, Auth, Storage, Assets…

    623 GitHub stars~2.2k tokensUpdated 11 days ago
    Backend & APIsAuto-check: notes
  • Operates Tianji Workers: create, test, deploy, invoke, schedule, pause and roll back them, plus manage their environment variables and shared modules.

    3.1k GitHub stars~1.1k tokensUpdated today
    Backend & APIsAuto-check passed
  • AI Model Nodejs

    TencentCloudBase/CloudBase-AI-Toolkit

    A skill your agent uses for Node.js backend AI via @cloudbase/node-sdk (=3.16.0) — cloud functions, CloudRun, Express/Koa/NestJS, serverless APIs, scheduled jobs, LLM proxies, agent orchestration.

    1.1k GitHub starsUsed in 2 repos~5k tokens
    Backend & APIsAuto-check passed
  • Qstash JS

    upstash/qstash-js

    Official

    Work with the QStash JavaScript/TypeScript SDK for serverless messaging, scheduling.

    269 GitHub stars~746 tokensUpdated 3 days ago
    Backend & APIsAuto-check passed

More from AI4Scientist/nano-scientist

All 74 skills in this repo
  • Formula Derivation

    AI4Scientist/nano-scientist

    Structures and derives research formulas when the user wants to 推导公式, build a theory line, organize assumptions, turn scattered equations into a coherent derivation, or rewrite theory notes into a…

    128 GitHub starsUsed in 5 repos~2.3k tokens
    Auto-check passed
  • Paper Compile

    AI4Scientist/nano-scientist

    Compile LaTeX paper to PDF, fix errors, and verify output. An agent skill from AI4Scientist/nano-scientist.

    128 GitHub starsUsed in 5 repos~2.5k tokens
    Auto-check: notes
  • Paper Figure

    AI4Scientist/nano-scientist

    Generate publication-quality figures and tables from experiment results.

    128 GitHub starsUsed in 5 repos~2.9k tokens
    Auto-check: notes
  • Paper Navigator

    AI4Scientist/nano-scientist

    Find and read academic papers: disambiguate queries, discover papers (search, citation traversal, recommendations, arXiv monitoring, trending, GitHub search), evaluate (TLDR, citations, code, SOTA)…

    128 GitHub stars~7.7k tokensUpdated 4 mo ago
    Auto-check: notes
  • Proof Writer

    AI4Scientist/nano-scientist

    Writes rigorous mathematical proofs for ML/AI theory. An agent skill from AI4Scientist/nano-scientist.

    128 GitHub starsUsed in 5 repos~1.9k tokens
    Auto-check passed
  • Ablation Planner

    AI4Scientist/nano-scientist

    A skill your agent uses when main results pass result-to-claim (claimsupported=yes or partial) and ablation studies are needed for paper submission.

    128 GitHub starsUsed in 4 repos~1.3k tokens
    Auto-check: notes

Categories

Questions about Run Experiment

What does Run Experiment do?

Deploy and run ML experiments on local, remote, Vast.ai, or Modal serverless GPU. Run Experiment is an agent skill from AI4Scientist/nano-scientist.ai, or Modal serverless GPU.

When should I use Run Experiment?

Run Experiment fits situations like: user says run experiment; deploy to server; needs to launch training jobs.

How do I install Run Experiment in Claude Code?

Run `npx skills add AI4Scientist/nano-scientist --skill run-experiment -a claude-code`. Or copy the skill folder (skills/run-experiment in AI4Scientist/nano-scientist) into .claude/skills/run-experiment in your project. Claude Code loads it when a task matches its description.

How do I install Run Experiment in Codex?

Run `npx skills add AI4Scientist/nano-scientist --skill run-experiment -a codex`. Or copy the skill folder (skills/run-experiment in AI4Scientist/nano-scientist) into .agents/skills/run-experiment in your project. Codex loads it when a task matches its description.

Can I use Run Experiment in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add AI4Scientist/nano-scientist --skill run-experiment -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/run-experiment, .gemini/skills/run-experiment, .github/skills/run-experiment and .opencode/skills/run-experiment in your project.

What does Run Experiment need to run?

Going by SKILL.md and its folder, Run Experiment needs the command-line tools its instructions call (ssh, python, git, modal, rsync and scp) and credentials named WANDB_API_KEY. Our summary lists: Python 3; Docker; A credential in WANDB_API_KEY; A credential in YOUR_KEY. Its frontmatter pre-approves these tools: Bash(*), Read, Grep, Glob, Edit, Write, Agent, Skill(serverless-modal).

Does Run Experiment access the network?

SKILL.md names 3 domains. In commands or code: wandb.ai; the agent is likely to contact it when it follows the instructions. As links in the text: cloud.vast.ai and modal.com. This is read from the text; nothing was executed.

Is Run Experiment safe to install?

Our automated static check of SKILL.md found notes only (pre-approves every shell command (allowed-tools: bash)), nothing it rates as a warning. It is not a guarantee. Review the folder before installing.

What licence does Run Experiment use?

No licence was found for Run Experiment or its repository. Without one, default copyright applies: ask the author before reusing or redistributing it.

How many tokens does Run Experiment use?

About 2.9k tokens (SKILL.md is roughly 12k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Run Experiment?

Skills that share tags, products or a category with Run Experiment: Arcgis To Portaljs (datopian/portaljs, 2.4k stars), AWS Serverless Eda (zxkane/aws-skills, 367 stars), Nubase (OtterMind/Nubase, 623 stars) and Tianji Worker Operations (msgbyte/tianji, 3.1k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Run Experiment?

AI4Scientist (a GitHub organization) maintains it in AI4Scientist/nano-scientist, which has 128 GitHub stars. The repository holds 74 skills in this directory. The repository was last updated on June 3, 2026.

Source: AI4Scientist/nano-scientist on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.