Agent skill

Experiment

by menkesu in menkesu/awesome-pm-skills

Designs, sanity-checks and reads out product experiments (A/B tests, holdouts, pre/post, geo and incrementality tests) and produces an experiment brief plus a readout with a ship / hold / replicate…

Custom licenceAuto-check passedMarketing & SEO

Install Experiment

skills CLI
$ npx skills add menkesu/awesome-pm-skills --skill experiment -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install menkesu/awesome-pm-skills experiment --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/menkesu/awesome-pm-skills.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/experiment .claude/skills/experiment && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
experiment
GitHub stars
433
Token cost
~5.2k tokens
SKILL.md length
2,973 words
Files
4 (incl. references, assets)
Skills in repo
25
Repo updated
First seen
Licence
Custom licence

At a glance

Designs, sanity-checks and reads out product experiments (A/B tests, holdouts, pre/post, geo and incrementality tests) and produces an experiment brief plus a readout with a ship / hold / replicate…

  • Works in 3 steps: Diagnose (ask before answering) → Pick the play → Run it
  • You ask should we A/B test this
  • SKILL.md covers When to use, Step 1: Diagnose (ask before…, Step 2: Pick the play and Step 3: Run it, plus 6 more sections
  • Instructions only: no scripts, shell commands, URLs or credentials in SKILL.md

What it does

Experiment is an agent skill from menkesu/awesome-pm-skills. Designs, sanity-checks and reads out product experiments (A/B tests, holdouts, pre/post, geo and incrementality tests) and produces an experiment brief plus a readout with a ship / hold / replicate call. Use when you ask 'should we A/B test this', 'how long should the test run', 'is this result real', 'sample size', 'stat sig', 'SRM', 'holdout', 'our win rate is low', or need to grade an existing test plan or readout. Draws on 191 insights from 71 Lenny's Podcast guests, led by Ronny Kohavi, Ramesh Johari and…

Its SKILL.md is about 5.2k tokens, which your agent loads only when the skill is triggered. The skill folder holds 5 other files, including reference files and assets (for example `references/frameworks.md` and `references/quotes.md`).

It sits in Marketing & SEO, covering A/B testing, Experimental design and Test generation. The repository describes itself as: Product management workflows grounded in Lenny's Podcast, for Claude Code, Codex, and Cursor.

When your agent uses it

  • You ask should we A/B test this
  • How long should the test run
  • Is this result real
  • Our win rate is low

Example prompts

  • “should we A/B test this”
  • “how long should the test run”
  • “is this result real”
  • “/experiment”

Workflow steps

3 steps, taken from the step headings in SKILL.md.

  1. Diagnose (ask before answering)
  2. Pick the play
  3. Run it

What it can do on your machine

Read from SKILL.md and the folder at commit c1b6e2d. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    No scripts in the folder and no shell commands in SKILL.md (its code samples are markdown).

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Experiment loads about 5.2k tokens when it runs, and up to ~13k if it reads all its reference files. Until then it costs about 135 tokens; SKILL.md has 2,973 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~135
When it runs · the whole SKILL.md, loaded when a task matches
~5.2k
With references · SKILL.md plus every file in references/, read only if the agent opens them
~13k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

Its licence (Custom licence) doesn't allow us to republish the file, so here is its outline and opening line. It has 2,973 words (~5,165 tokens).

“Turn an idea into a pre-registered experiment, catch the ways it can lie to you, and end with a decision you can defend. Built from 191 insights from 71 Lenny's Podcast guests. Lead team: Ronny Kohavi, Ramesh Johari, Archie Abrams.”

— opening of SKILL.md by menkesu, Custom licence
name
experiment

Read the full SKILL.md on GitHub

Files

SKILL.md and 3 other files (references, assets) in skills/experiment of menkesu/awesome-pm-skills.

  • SKILL.md
  • assets/card.png
  • references/frameworks.md
  • references/quotes.md

Open the folder on GitHubat commit c1b6e2d

Compare with similar skills

Experiment next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Experiment compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Experiment this skillmenkesu/awesome-pm-skills433—~5.2kAutomated safety check: PassCustom licence
Ab Test ArchitectLeoYeAI/openclaw-master-skills2.2k—~5kAutomated safety check: PassMIT
Ab Test Plannermohitagw15856/pm-claude-skills1.4k—~2kAutomated safety check: PassMIT
Ad Test Designeraaron-he-zhu/aaron-marketing-skills2.9k2 repos~2.8kAutomated safety check: PassApache-2.0
Ab Test Analyzeririnabuht12-oss/marketing-skills4k—~1.4kAutomated safety check: PassNone
Define Hypothesisproduct-on-purpose/pm-skills715—~966Automated safety check: PassApache-2.0

Similar skills

  • Ab Test Architect

    LeoYeAI/openclaw-master-skills

    Plan, prioritize, and design rigorous A/B tests using the Test Velocity Method.

    2.2k GitHub stars~5k tokensUpdated 2 mo ago
    Marketing & SEOAuto-check passed
  • Ab Test Planner

    mohitagw15856/pm-claude-skills

    Design statistically rigorous A/B tests for product features, UI changes, onboarding flows, and pricing experiments.

    1.4k GitHub stars~2k tokensUpdated yesterday
    Marketing & SEOAuto-check passed
  • Ad Test Designer

    aaron-he-zhu/aaron-marketing-skills

    A skill your agent uses when the user asks to "design an A/B test", "set up a creative/landing test", "run an incrementality test", or "is this result statistically and practically material?"…

    2.9k GitHub starsUsed in 2 repos~2.8k tokens
    Marketing & SEOAuto-check passed
  • Ab Test Analyzer

    irinabuht12-oss/marketing-skills

    Statistical significance calculator for A/B test results with sample size requirements, segment breakdowns, and hypothesis generation.

    4k GitHub stars~1.4k tokensUpdated 15 days ago
    Marketing & SEOAuto-check passed
  • Define Hypothesis

    product-on-purpose/pm-skills

    Defines a testable hypothesis with clear success metrics and a validation approach.

    715 GitHub stars~966 tokensUpdated yesterday
    Marketing & SEOAuto-check passed
  • A B Test Design

    Owl-Listener/designer-skills

    Design an A/B experiment — hypothesis, variants, primary metric, and sample size.

    2.9k GitHub starsUsed in 1 repo~472 tokens
    Marketing & SEOAuto-check passed

More from menkesu/awesome-pm-skills

All 25 skills in this repo
  • Craft Review

    menkesu/awesome-pm-skills

    Runs a product quality review on a journey, feature, landing page, onboarding flow, or AI/agent surface and returns a friction log, a scored critical-journey review, and a taste bar your team can…

    433 GitHub stars~4.6k tokensUpdated 2 days ago
    Auto-check passed
  • Customer Interviews

    menkesu/awesome-pm-skills

    Designs customer interviews and turns the transcripts into decisions.

    433 GitHub stars~4.4k tokensUpdated 2 days ago
    Auto-check passed
  • Decide

    menkesu/awesome-pm-skills

    Turns a hard product or career decision into a one-page decision memo with the eigenquestion, options, pre-mortem with kill criteria, reversibility rating, a named decider and a review date.

    433 GitHub stars~4.9k tokensUpdated 2 days ago
    Auto-check passed
  • Eval Plan

    menkesu/awesome-pm-skills

    Builds an eval plan for an AI feature - error analysis on real traces, failure-mode ranking, code checks and binary LLM judges validated against human labels, CI tests, production monitoring, and a…

    433 GitHub stars~4.5k tokensUpdated 2 days ago
    Auto-check passed
  • Exec Comms

    menkesu/awesome-pm-skills

    Rewrite an exec update, memo, strategy doc, deck outline, Slack post or talk so it lands with senior people, and get a graded score plus the top 3 fixes.

    433 GitHub stars~4.4k tokensUpdated 2 days ago
    Auto-check passed
  • Hard Conversations

    menkesu/awesome-pm-skills

    Writes you a ready-to-use script for a feedback or hard conversation, then rehearses the pushback.

    433 GitHub stars~4.7k tokensUpdated 2 days ago
    Auto-check passed

Questions about Experiment

What does Experiment do?

Designs, sanity-checks and reads out product experiments (A/B tests, holdouts, pre/post, geo and incrementality tests) and produces an experiment brief plus a readout with a ship / hold / replicate…. Experiment is an agent skill from menkesu/awesome-pm-skills. Designs, sanity-checks and reads out product experiments (A/B tests, holdouts, pre/post, geo and incrementality tests) and produces an experiment brief plus a readout with a ship / hold / replicate call.

When should I use Experiment?

Experiment fits situations like: you ask should we A/B test this; how long should the test run; is this result real; our win rate is low.

How do I install Experiment in Claude Code?

Run `npx skills add menkesu/awesome-pm-skills --skill experiment -a claude-code`. Or copy the skill folder (skills/experiment in menkesu/awesome-pm-skills) into .claude/skills/experiment in your project. Claude Code loads it when a task matches its description.

How do I install Experiment in Codex?

Run `npx skills add menkesu/awesome-pm-skills --skill experiment -a codex`. Or copy the skill folder (skills/experiment in menkesu/awesome-pm-skills) into .agents/skills/experiment in your project. Codex loads it when a task matches its description.

Can I use Experiment in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add menkesu/awesome-pm-skills --skill experiment -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/experiment, .gemini/skills/experiment, .github/skills/experiment and .opencode/skills/experiment in your project.

What does Experiment need to run?

SKILL.md names no scripts, command-line tools or credentials: Experiment is instructions for the agent only.

Does Experiment access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is Experiment safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Experiment use?

Experiment has a licence file (the repository's licence) that doesn't match a standard licence. Read it on GitHub before reusing the skill.

How many tokens does Experiment use?

About 5.2k tokens (SKILL.md is roughly 21k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full. Its references folder adds about 8.2k tokens, read only when the agent opens those files.

What are the alternatives to Experiment?

Skills that share tags, products or a category with Experiment: Ab Test Architect (LeoYeAI/openclaw-master-skills, 2.2k stars), Ab Test Planner (mohitagw15856/pm-claude-skills, 1.4k stars), Ad Test Designer (aaron-he-zhu/aaron-marketing-skills, 2.9k stars) and Ab Test Analyzer (irinabuht12-oss/marketing-skills, 4k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Experiment?

menkesu (a GitHub user) maintains it in menkesu/awesome-pm-skills, which has 433 GitHub stars. The repository holds 25 skills in this directory. The repository was last updated on October 6, 2026.

Source: menkesu/awesome-pm-skills on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.