Agent skill

Monitor Experiment

by AI4Scientist in AI4Scientist/nano-scientist

Monitor running experiments, check progress, collect results.

No licenceAuto-check passedAgent Workflows

Install Monitor Experiment

skills CLI
$ npx skills add AI4Scientist/nano-scientist --skill monitor-experiment -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install AI4Scientist/nano-scientist monitor-experiment --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/AI4Scientist/nano-scientist.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/monitor-experiment .claude/skills/monitor-experiment && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
monitor-experiment
GitHub stars
128
Used in
5 other repos
Token cost
~1.1k tokens
SKILL.md length
361 words
Files
1
Skills in repo
75
Repo updated
First seen
Licence
None found

At a glance

Monitor running experiments, check progress, collect results.

  • Works in 7 steps: Check What's Running → Collect Output from Each Screen → Check for JSON Result Files → …
  • User says check results
  • SKILL.md covers Workflow and Key Rules
  • Calls ssh and modal; reaches wandb.ai

What it does

Monitor Experiment is an agent skill from AI4Scientist/nano-scientist. Monitor running experiments, check progress, collect results. Use when user says "check results", "is it done", "monitor", or wants experiment output.

Its SKILL.md is about 1.1k tokens, which your agent loads only when the skill is triggered. It is a single SKILL.md file with no bundled scripts.

It sits in Agent Workflows. It works with Weights & Biases. The repository describes itself as: An autonomous research agent that turns a topic into a peer-reviewed technical report.

When your agent uses it

  • User says check results
  • Wants experiment output

Example prompts

  • “check results”
  • “is it done”
  • “monitor”
  • “/monitor-experiment”

Requirements

  • Python 3
  • Pre-approved tools (allowed-tools): Bash(ssh *), Bash(echo *), Read, Write, Edit

Workflow steps

7 steps, taken from the step headings in SKILL.md.

  1. Check What's Running
  2. Collect Output from Each Screen
  3. Check for JSON Result Files
  4. 5: Pull W&B Metrics (when wandb: true in CLAUDE.md)
  5. Summarize Results
  6. Interpret
  7. Feishu Notification (if configured)

What it can do on your machine

Read from SKILL.md and the folder at commit 7132192. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves these tools, so the agent can use them without asking each time:

    • Bash(ssh *)
    • Bash(echo *)
    • Read
    • Write
    • Edit

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Shell commands in SKILL.md call:

    • ssh
    • modal

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    Hosts in commands or code, which the agent is likely to contact:

    • wandb.ai

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Monitor Experiment loads about 1.1k tokens when it runs. Until then it costs about 42 tokens; SKILL.md has 361 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~42
When it runs · the whole SKILL.md, loaded when a task matches
~1.1k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

Without a licence we can't republish the file, so here is its outline and opening line. It has 361 words (~1,109 tokens).

name
monitor-experiment
allowed-tools
Bash(ssh *), Bash(echo *), Read, Write, Edit
argument-hint
server-alias or screen-name

Read the full SKILL.md on GitHub

Files

Just SKILL.md in skills/monitor-experiment of AI4Scientist/nano-scientist.

Open the folder on GitHubat commit 7132192

Used in 5 other repositories

We found 11 copies of this SKILL.md (exact, near-identical or edited) in other folders, from 5 other GitHub owners. This page covers the copy in AI4Scientist/nano-scientist, which our catalogue first saw on October 7, 2026.

Compare with similar skills

Monitor Experiment next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Monitor Experiment compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Monitor Experiment this skillAI4Scientist/nano-scientist1285 repos~1.1kAutomated safety check: PassNone
Hermes Atropos EnvironmentsTommy-yw/RunbookHermes546—~3.3kAutomated safety check: PassMIT
Sft Launchopen-thoughts/OpenThoughts-Agent301—~2.9kAutomated safety check: PassApache-2.0
DashboardLegoX/Lego-RL108—~4.2kAutomated safety check: PassApache-2.0
Azure Mgmt Weightsandbiases Dotnetmicrosoft/skills3.1k6 repos~2.8kAutomated safety check: PassMIT
Crud Archive Runopen-thoughts/OpenThoughts-Agent301—~1.2kAutomated safety check: PassApache-2.0

Similar skills

  • Hermes Atropos Environments

    Tommy-yw/RunbookHermes

    Build, test, and debug Hermes Agent RL environments for Atropos training.

    546 GitHub stars~3.3k tokensUpdated 4 mo ago
    AI & LLM EngineeringAuto-check passed
  • Sft Launch

    open-thoughts/OpenThoughts-Agent

    Launch SFT via python -m hpc.launch --jobtype sft on any cluster (JSC Jupiter GH200, CINECA Leonardo A100, TACC Vista GH200), with EITHER backend — LLaMA-Factory (default) or axolotl (--sftbackend…

    301 GitHub stars~2.9k tokensUpdated 9 days ago
    AI & LLM EngineeringAuto-check passed
  • Dashboard

    LegoX/Lego-RL

    Bring up the Lego-RL training dashboard (webui/) on whatever machine you are on, adapting to that box's layout instead of assuming this repo's paths.

    108 GitHub stars~4.2k tokensUpdated today
    AI & LLM EngineeringAuto-check passed
  • Official

    Azure Weights & Biases SDK for .NET. An agent skill from microsoft/skills.

    3.1k GitHub starsUsed in 6 repos~2.8k tokens
    DevOps & CloudAuto-check passed
  • Crud Archive Run

    open-thoughts/OpenThoughts-Agent

    Durably ARCHIVE everything informative from a finished run / experiment before it's cleaned up or its cluster artifacts age out — ALL Harbor tracejobs (raw per-trial traces), ALL ray logs, ALL…

    301 GitHub stars~1.2k tokensUpdated 9 days ago
    AI & LLM EngineeringAuto-check passed
  • Onboard Marin

    marin-community/marin

    Verify or complete a new internal Marin developer's local setup and access to GitHub, GCP, Iris, Weights & Biases, Hugging Face, and optional CoreWeave storage.

    3.9k GitHub stars~1.1k tokensUpdated today
    AI & LLM EngineeringAuto-check passed

More from AI4Scientist/nano-scientist

All 75 skills in this repo
  • Formula Derivation

    AI4Scientist/nano-scientist

    Structures and derives research formulas when the user wants to 推导公式, build a theory line, organize assumptions, turn scattered equations into a coherent derivation, or rewrite theory notes into a…

    128 GitHub starsUsed in 6 repos~2.3k tokens
    Auto-check passed
  • Paper Figure

    AI4Scientist/nano-scientist

    Generate publication-quality figures and tables from experiment results.

    128 GitHub starsUsed in 6 repos~2.9k tokens
    Auto-check: notes
  • Proof Writer

    AI4Scientist/nano-scientist

    Writes rigorous mathematical proofs for ML/AI theory. An agent skill from AI4Scientist/nano-scientist.

    128 GitHub starsUsed in 6 repos~1.9k tokens
    Auto-check passed
  • Ablation Planner

    AI4Scientist/nano-scientist

    A skill your agent uses when main results pass result-to-claim (claimsupported=yes or partial) and ablation studies are needed for paper submission.

    128 GitHub starsUsed in 5 repos~1.3k tokens
    Auto-check: notes
  • Paper Navigator

    AI4Scientist/nano-scientist

    Find and read academic papers: disambiguate queries, discover papers (search, citation traversal, recommendations, arXiv monitoring, trending, GitHub search), evaluate (TLDR, citations, code, SOTA)…

    128 GitHub stars~7.7k tokensUpdated 4 mo ago
    Auto-check: notes
  • Arxiv

    AI4Scientist/nano-scientist

    Search, download, and summarize academic papers from arXiv. An agent skill from AI4Scientist/nano-scientist.

    128 GitHub starsUsed in 5 repos~2.1k tokens
    Auto-check: notes

Categories

Questions about Monitor Experiment

What does Monitor Experiment do?

Monitor running experiments, check progress, collect results. Monitor Experiment is an agent skill from AI4Scientist/nano-scientist. Monitor running experiments, check progress, collect results.

When should I use Monitor Experiment?

Monitor Experiment fits situations like: user says check results; wants experiment output.

How do I install Monitor Experiment in Claude Code?

Run `npx skills add AI4Scientist/nano-scientist --skill monitor-experiment -a claude-code`. Or copy the skill folder (skills/monitor-experiment in AI4Scientist/nano-scientist) into .claude/skills/monitor-experiment in your project. Claude Code loads it when a task matches its description.

How do I install Monitor Experiment in Codex?

Run `npx skills add AI4Scientist/nano-scientist --skill monitor-experiment -a codex`. Or copy the skill folder (skills/monitor-experiment in AI4Scientist/nano-scientist) into .agents/skills/monitor-experiment in your project. Codex loads it when a task matches its description.

Can I use Monitor Experiment in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add AI4Scientist/nano-scientist --skill monitor-experiment -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/monitor-experiment, .gemini/skills/monitor-experiment, .github/skills/monitor-experiment and .opencode/skills/monitor-experiment in your project.

What does Monitor Experiment need to run?

Going by SKILL.md and its folder, Monitor Experiment needs the command-line tools its instructions call (ssh and modal). Our summary lists: Python 3. Its frontmatter pre-approves these tools: Bash(ssh *), Bash(echo *), Read, Write, Edit.

Does Monitor Experiment access the network?

SKILL.md names 1 domain. In commands or code: wandb.ai; the agent is likely to contact it when it follows the instructions. This is read from the text; nothing was executed.

Is Monitor Experiment safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Monitor Experiment use?

No licence was found for Monitor Experiment or its repository. Without one, default copyright applies: ask the author before reusing or redistributing it.

How many tokens does Monitor Experiment use?

About 1.1k tokens (SKILL.md is roughly 4.4k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Monitor Experiment?

Skills that share tags, products or a category with Monitor Experiment: Hermes Atropos Environments (Tommy-yw/RunbookHermes, 546 stars), Sft Launch (open-thoughts/OpenThoughts-Agent, 301 stars), Dashboard (LegoX/Lego-RL, 108 stars) and Azure Mgmt Weightsandbiases Dotnet (microsoft/skills, 3.1k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Monitor Experiment?

AI4Scientist (a GitHub organization) maintains it in AI4Scientist/nano-scientist, which has 128 GitHub stars. The repository holds 75 skills in this directory. The repository was last updated on June 3, 2026.

Source: AI4Scientist/nano-scientist on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.