Agent skill

Organize Experiments

by marin-community in marin-community/marin

Harvest experiment issue reports and curate docs/reports/index.md only when explicitly requested.

Apache-2.0Auto-check passed

Install Organize Experiments

skills CLI
$ npx skills add marin-community/marin --skill organize-experiments -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install marin-community/marin organize-experiments --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/marin-community/marin.git skills-src && mkdir -p .claude/skills && cp -r skills-src/.agents/skills/organize-experiments .claude/skills/organize-experiments && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
organize-experiments
GitHub stars
3.9k
Token cost
~645 tokens
SKILL.md length
300 words
Files
1
Skills in repo
41
Repo updated
First seen
Licence
Apache-2.0

At a glance

Harvest experiment issue reports and curate docs/reports/index.md only when explicitly requested.

  • Works in 6 steps: Run uv run… → Open the diff for docs/reports/index.md… → For each experiment → …
  • SKILL.md covers Overview, Prerequisites, Guidelines for Humans and Rules for Agents, plus 2 more sections
  • Calls uv and rg; reaches wandb.ai and img.shields.io

What it does

Organize Experiments is an agent skill from marin-community/marin. Harvest experiment issue reports and curate docs/reports/index.md only when explicitly requested.

Its SKILL.md is about 650 tokens, which your agent loads only when the skill is triggered. It is a single SKILL.md file with no bundled scripts.

It works with Weights & Biases. The repository describes itself as: Open-source framework for the research and development of foundation models. The licence is Apache-2.0.

Example prompts

  • “/organize-experiments”

Workflow steps

6 steps, taken from the first numbered list in SKILL.md.

  1. Run uv run scripts/pm/itemize_experiment_issues.py to append new experiments.
  2. Open the diff for docs/reports/index.md and identify the additions in ## Uncategorized.
  3. For each experiment
  4. If an experiment truly does not fit, leave it under ## Uncategorized and add a note explaining why.
  5. Proofread for duplicate bullets, broken Markdown, and consistent title casing.
  6. Commit the curated report, open a PR, or, if an interactive agent, signal to the user to look.

What it can do on your machine

Read from SKILL.md and the folder at commit 61bb85c. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Shell commands in SKILL.md call:

    • uv
    • rg

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    Hosts in commands or code, which the agent is likely to contact:

    • wandb.ai
    • img.shields.io

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Organize Experiments loads about 645 tokens when it runs. Until then it costs about 30 tokens; SKILL.md has 300 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~30
When it runs · the whole SKILL.md, loaded when a task matches
~645

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from marin-community/marin at commit 61bb85c, republished under its Apache-2.0 licence (© marin-community). 300 words, ~645 tokens.

Download SKILL.mdSave it as .claude/skills/organize-experiments/SKILL.md (or your agent's skills folder).
name
organize-experiments
description
Harvest experiment issue reports and curate docs/reports/index.md only when explicitly requested.

Skill: Organize Experiment Reports

Overview

Curate docs/reports/index.md after new experiment issues are harvested: fold fresh entries into the right sections, refresh links, and leave ## Uncategorized empty.

Prerequisites

  • Local checkout of the marin repository with write access.
  • Ability to run uv commands.
  • Familiarity with the existing experiment categories in docs/reports/index.md.

Guidelines for Humans

Standard Workflow
  1. Run uv run scripts/pm/itemize_experiment_issues.py to append new experiments.
  2. Open the diff for docs/reports/index.md and identify the additions in ## Uncategorized.
  3. For each experiment:
    • Match it to an existing section (e.g., Training and Performance, Data Experiments).
    • Merge new WandB links into the canonical entry in that section.
    • Remove the placeholder entry from ## Uncategorized.
  4. If an experiment truly does not fit, leave it under ## Uncategorized and add a note explaining why.
  5. Proofread for duplicate bullets, broken Markdown, and consistent title casing.
  6. Commit the curated report, open a PR, or, if an interactive agent, signal to the user to look.
  • Prefer direct https://wandb.ai/... links when available; keep legacy api.wandb.ai links only if the direct share link does not exist.
  • Do not add Data Browser links. The service was retired in #8624; the script collects only WandB URLs.
  • Preserve access tokens embedded in links.

Rules for Agents

  • NEVER delete existing sections or conclusions.
  • Keep badge styling consistent: [![#NNN](https://img.shields.io/...)] next to each experiment title.
  • When merging new content, update the canonical entry instead of duplicating it.
  • Leave ## Uncategorized empty whenever possible; a single placeholder sentence is fine.
  • Ask for guidance if an experiment does not map cleanly to known categories.

Validation

  • rg "^- " docs/reports/index.md to confirm bullets exist only under curated sections, not under ## Uncategorized.
  • Re-run uv run scripts/pm/itemize_experiment_issues.py if unsure that all experiments were captured.
  • Optional: markdownlint docs/reports/index.md to catch formatting drift.

See Also

  • scripts/pm/itemize_experiment_issues.py for source data generation.
  • .agents/skills/archive-experiments/SKILL.md for retiring legacy experiments.

© marin-community, Apache-2.0. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

Just SKILL.md in .agents/skills/organize-experiments of marin-community/marin.

Open the folder on GitHubat commit 61bb85c

Compare with similar skills

Organize Experiments next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Organize Experiments compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Organize Experiments this skillmarin-community/marin3.9k—~645Automated safety check: PassApache-2.0
Marimo Batchkoaning/gitcharts1451 repos~819Automated safety check: NotesNone
Weights & Biases Experiment TrackingOrchestra-Research/AI-Research-SKILLs13k9 repos~3.1kAutomated safety check: PassMIT
Comparefcakyon/phd-skills415—~1.2kAutomated safety check: PassMIT
ML Experiment IterationLeeroo-AI/superml195—~4.8kAutomated safety check: PassApache-2.0
LaminDB Biological Data Managementdavila7/claude-code-templates32k12 repos~3.6kAutomated safety check: PassMIT

Similar skills

  • Marimo Batch

    koaning/gitcharts

    An opintionated skill to prepare a marimo notebook to make it ready for a scheduled run.

    145 GitHub starsUsed in 1 repo~819 tokens
    Data & AnalyticsAuto-check: notes
  • Weights & Biases Experiment Tracking

    Orchestra-Research/AI-Research-SKILLs

    Guides an agent through tracking ML experiments with W&B: run logging, config capture, hyperparameter sweeps, artifacts and a model registry.

    13k GitHub starsUsed in 9 repos~3.1k tokens
    DevOps & CloudAuto-check passed
  • Compare

    fcakyon/phd-skills

    Same-epoch comparison of training runs across wandb, neptune, tensorboard, or mlflow.

    415 GitHub stars~1.2k tokensUpdated 23 days ago
    Research & ScienceAuto-check passed
  • ML Experiment Iteration

    Leeroo-AI/superml

    Produces ranked, evidence-grounded next steps when an ML experiment has stalled, drawing on a Leeroopedia knowledge base or on fetched docs and issues.

    195 GitHub stars~4.8k tokensUpdated 6 mo ago
    AI & LLM EngineeringAuto-check passed
  • LaminDB Biological Data Management

    davila7/claude-code-templates

    Manages biological datasets with LaminDB: versioned artifacts, run lineage, ontology-based annotation, schema validation and links to workflow managers and ML tools.

    32k GitHub starsUsed in 12 repos~3.6k tokens
    Research & ScienceAuto-check passed
  • Dashboard

    LegoX/Lego-RL

    Bring up the Lego-RL training dashboard (webui/) on whatever machine you are on, adapting to that box's layout instead of assuming this repo's paths.

    111 GitHub stars~4.2k tokensUpdated yesterday
    AI & LLM EngineeringAuto-check passed

More from marin-community/marin

All 41 skills in this repo
  • Noslop

    marin-community/marin

    Deslop, simplify, or review low-value tests and prose only when explicitly requested for a branch or diff.

    3.9k GitHub stars~1.1k tokensUpdated today
    Auto-check passed
  • Use Iris

    marin-community/marin

    Use Iris to submit, inspect, debug, monitor, or recover jobs and tasks; diagnose scheduling and federation; deploy controllers; or reserve dev GPUs and TPUs.

    3.9k GitHub stars~745 tokensUpdated today
    Auto-check passed
  • Launch Rl

    marin-community/marin

    Define, validate, submit, or restart a Marin SkyRL experiment through its artifact main.

    3.9k GitHub stars~894 tokensUpdated today
    Auto-check passed
  • Marina Applet

    marin-community/marin

    Build, validate, publish, update, inspect, query, roll back, or archive a dynamic Marina applet.

    3.9k GitHub stars~2.3k tokensUpdated today
    Auto-check passed
  • Query Finelog

    marin-community/marin

    Query Finelog logs and telemetry for Iris tasks, workers, profiles, training, vLLM, and cross-cluster forwarding.

    3.9k GitHub stars~1.1k tokensUpdated today
    Auto-check passed
  • Trace Pulumi Diff

    marin-community/marin

    Run a read-only preview for a specified Marin infra/pulumi stack and trace each pending resource change to merged pull requests since its latest successful update when that update records a clean…

    3.9k GitHub stars~663 tokensUpdated today
    Auto-check passed

Questions about Organize Experiments

What does Organize Experiments do?

Harvest experiment issue reports and curate docs/reports/index.md only when explicitly requested. Organize Experiments is an agent skill from marin-community/marin.md only when explicitly requested.

How do I install Organize Experiments in Claude Code?

Run `npx skills add marin-community/marin --skill organize-experiments -a claude-code`. Or copy the skill folder (.agents/skills/organize-experiments in marin-community/marin) into .claude/skills/organize-experiments in your project. Claude Code loads it when a task matches its description.

How do I install Organize Experiments in Codex?

Run `npx skills add marin-community/marin --skill organize-experiments -a codex`. Or copy the skill folder (.agents/skills/organize-experiments in marin-community/marin) into .agents/skills/organize-experiments in your project. Codex loads it when a task matches its description.

Can I use Organize Experiments in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add marin-community/marin --skill organize-experiments -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/organize-experiments, .gemini/skills/organize-experiments, .github/skills/organize-experiments and .opencode/skills/organize-experiments in your project.

What does Organize Experiments need to run?

Going by SKILL.md and its folder, Organize Experiments needs the command-line tools its instructions call (uv and rg).

Does Organize Experiments access the network?

SKILL.md names 2 domains. In commands or code: wandb.ai and img.shields.io; the agent is likely to contact these when it follows the instructions. This is read from the text; nothing was executed.

Is Organize Experiments safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Organize Experiments use?

Organize Experiments is published under the Apache-2.0 licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Organize Experiments use?

About 645 tokens (SKILL.md is roughly 2.6k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Organize Experiments?

Skills that share tags, products or a category with Organize Experiments: Marimo Batch (koaning/gitcharts, 145 stars), Weights & Biases Experiment Tracking (Orchestra-Research/AI-Research-SKILLs, 13k stars), Compare (fcakyon/phd-skills, 415 stars) and ML Experiment Iteration (Leeroo-AI/superml, 195 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Organize Experiments?

marin-community (a GitHub organization) maintains it in marin-community/marin, which has 3,920 GitHub stars. The repository holds 41 skills in this directory. The repository was last updated on October 9, 2026.

Source: marin-community/marin on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.