Agent skill

Experiment Design Kit

by gtmagents in gtmagents/gtm-agents

Toolkit for structuring hypotheses, variants, guardrails, and measurement plans.

Apache-2.0Auto-check passedAI & LLM Engineering

Install Experiment Design Kit

skills CLI
$ npx skills add gtmagents/gtm-agents --skill experiment-design-kit -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install gtmagents/gtm-agents experiment-design-kit --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/gtmagents/gtm-agents.git skills-src && mkdir -p .claude/skills && cp -r skills-src/plugins/growth-experiments/skills/experiment-design-kit .claude/skills/experiment-design-kit && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
experiment-design-kit
GitHub stars
414
Used in
1 other repo
Token cost
~837 tokens
SKILL.md length
352 words
Files
1
Skills in repo
121
Repo updated
First seen
Licence
Apache-2.0

At a glance

Toolkit for structuring hypotheses, variants, guardrails, and measurement plans.

  • Works in 5 steps: Problem Framing – define user problem,… → Hypothesis Structure – "If we do X for Y… → Measurement Plan – primary metric,… → …
  • Tasks that involve Experimental design
  • SKILL.md covers When to Use, Framework, Templates and Tips, plus 3 more sections
  • Instructions only: no scripts, shell commands, URLs or credentials in SKILL.md

What it does

Experiment Design Kit is an agent skill from gtmagents/gtm-agents. Toolkit for structuring hypotheses, variants, guardrails, and measurement plans.

Its SKILL.md is about 840 tokens, which your agent loads only when the skill is triggered. It is a single SKILL.md file with no bundled scripts.

It sits in AI & LLM Engineering, covering Experimental design, LLM guardrails and Go-to-market strategy. The repository describes itself as: Production-ready collection of GTM agents and specialized skills for claude code. Covers sales, marketing, customer success, and revenue operations workflows. The licence is Apache-2.0.

When your agent uses it

  • Tasks that involve Experimental design
  • Tasks that involve LLM guardrails
  • Tasks that involve Go-to-market strategy

Example prompts

  • “/experiment-design-kit”

Workflow steps

5 steps, taken from the first numbered list in SKILL.md.

  1. Problem Framing – define user problem, business impact, and north-star metric.
  2. Hypothesis Structure – "If we do X for Y persona, we expect Z change" with assumptions.
  3. Measurement Plan – primary metric, guardrails, min detectable effect, power calc.
  4. Variant Strategy – control definition, variant catalog, targeting, and exclusion rules.
  5. Operational Plan – owners, timeline, dependencies, QA/rollback steps.

What it can do on your machine

Read from SKILL.md and the folder at commit 78e0419. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    No scripts in the folder and no shell commands in SKILL.md.

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Experiment Design Kit loads about 837 tokens when it runs. Until then it costs about 26 tokens; SKILL.md has 352 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~26
When it runs · the whole SKILL.md, loaded when a task matches
~837

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from gtmagents/gtm-agents at commit 78e0419, republished under its Apache-2.0 licence (© gtmagents). 352 words, ~837 tokens.

Download SKILL.mdSave it as .claude/skills/experiment-design-kit/SKILL.md (or your agent's skills folder).
name
experiment-design-kit
description
Toolkit for structuring hypotheses, variants, guardrails, and measurement plans.

Experiment Design Kit Skill

When to Use

  • Translating raw ideas into testable hypotheses with clear success metrics.
  • Ensuring experiment briefs include guardrails, instrumentation, and rollout details.
  • Coaching pods on best practices for multi-variant or multi-surface tests.

Framework

  1. Problem Framing – define user problem, business impact, and north-star metric.
  2. Hypothesis Structure – "If we do X for Y persona, we expect Z change" with assumptions.
  3. Measurement Plan – primary metric, guardrails, min detectable effect, power calc.
  4. Variant Strategy – control definition, variant catalog, targeting, and exclusion rules.
  5. Operational Plan – owners, timeline, dependencies, QA/rollback steps.

Templates

  • Experiment brief (context, hypothesis, design, metrics, launch checklist).
  • Guardrail register with thresholds + alerting rules.
  • Variant matrix for surfaces, messaging, and states.
  • GTM Agents Growth Backlog Board – capture idea → sizing → prioritization scoring (ICE/RICE) @puerto/README.md#183-212.
  • Weekly Experiment Packet – includes KPI guardrails, qualitative notes, and next bets for Marketing Director + Sales Director.
  • Rollback Playbook – pre-built checklist tied to lifecycle-mapping rip-cord procedures.

Tips

  • Pressure-test hypotheses with counter-metrics to avoid local optima.
  • Document data constraints early to avoid rework during build.
  • Pair with guardrail-scorecard to ensure sign-off before launch.
  • Apply GTM Agents cadence: Monday backlog groom, Wednesday build review, Friday learnings sync.
  • Require KPI guardrails per stage (activation, engagement, monetization) before authorizing build.
  • If a test risks Sales velocity, include Sales Director in approval routing per GTM Agents governance.
Show full SKILL.md (133 more words)Show less

GTM Agents Experiment Operating Model

  1. Backlog Intake – ideas flow from GTM pods; Growth Marketer tags theme, objective, expected impact.
  2. Prioritization – score with RICE + qualitative "strategic fit" modifier; surface top 3 bets weekly.
  3. Design & Instrumentation – reference Serena/Context7 to patch code + confirm documentation.
  4. Launch & Monitor – use guardrail-scorecard to watch leading indicators (churn, complaints, latency).
  5. Learning Loop – run Sequential Thinking retro; document hypothesis, result, decision, follow-up in backlog card.

KPI Guardrails (GTM Agents Reference)

  • Activation rate change must stay within ±3% of baseline for Tier-1 segments.
  • Revenue per visitor cannot drop more than 2% for more than 48h.
  • Support tickets tied to experiment variant must remain <5% of total volume.

Weekly Experiment Packet Outline

Week Ending: <Date>

1. Portfolio Snapshot – tests live, status, KPI trend (guardrail vs actual)
2. Key Wins – hypothesis, uplift, next action (ship, iterate, expand)
3. Guardrail Alerts – what tripped, mitigation taken (rollback? scope adjust?)
4. Pipeline Impact – SQLs, ARR influenced, notable customer anecdotes
5. Upcoming Launches – dependencies, owners, open questions

Share packet with Growth, Marketing Director, Sales Director, and RevOps to mirror GTM Agents's cross-functional communication rhythm.


© gtmagents, Apache-2.0. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

Just SKILL.md in plugins/growth-experiments/skills/experiment-design-kit of gtmagents/gtm-agents.

Open the folder on GitHubat commit 78e0419

Used in 1 other repository

We found 1 copy of this SKILL.md (exact, near-identical or edited) in other folders, from 1 other GitHub owner. This page covers the copy in gtmagents/gtm-agents, which our catalogue first saw on October 7, 2026.

Compare with similar skills

Experiment Design Kit next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Experiment Design Kit compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Experiment Design Kit this skillgtmagents/gtm-agents4141 repos~837Automated safety check: PassApache-2.0
Data Scientistnagisanzenin/production-grade181—~3kAutomated safety check: PassNone
Phi Prompt Guardaipoch/medical-research-skills1.9k—~4.2kAutomated safety check: PassMIT
Acm Conference On Fairness Accountability And Transparencyfranklee16/academic-research-skills2231 repos~2kAutomated safety check: PassNone
Aisafetyhotwuyoscar/AISafetyHot-Hub827—~1.4kAutomated safety check: PassCustom licence
ObliteratusRedWoodOG/Hermes-Desktop1775 repos~3.8kAutomated safety check: PassMIT

Similar skills

  • Data Scientist

    nagisanzenin/production-grade

    [production-grade internal] Optimizes AI/ML/LLM usage when you need model selection, prompt engineering, cost reduction, or experiment design.

    181 GitHub stars~3k tokensUpdated 1 mo ago
    AI & LLM EngineeringAuto-check passed
  • Phi Prompt Guard

    aipoch/medical-research-skills

    Runtime, prompt-time behavioral guardrail that helps reduce PHI exposure in LLM-assisted workflows by detecting PHI-bearing prompts, avoiding unsafe tool actions that would pull more PHI in, and…

    1.9k GitHub stars~4.2k tokensUpdated 24 days ago
    AI & LLM EngineeringAuto-check passed
  • A skill your agent uses when targeting ACM Conference on Fairness, Accountability, and Transparency (FAccT) or deciding whether a computer-science manuscript fits this venue.

    223 GitHub starsUsed in 1 repo~2k tokens
    AI & LLM EngineeringAuto-check passed
  • Aisafetyhot

    wuyoscar/AISafetyHot-Hub

    Query AI Safety HOT news, research papers, incidents, hot topics, and daily/weekly/monthly reports through its public read-only MCP service.

    827 GitHub stars~1.4k tokensUpdated today
    AI & LLM EngineeringAuto-check passed
  • Obliteratus

    RedWoodOG/Hermes-Desktop

    Remove refusal behaviors from open-weight LLMs using OBLITERATUS — mechanistic interpretability techniques (diff-in-means, SVD, whitened SVD, LEACE, SAE decomposition, etc.) to excise guardrails…

    177 GitHub starsUsed in 5 repos~3.8k tokens
    AI & LLM EngineeringAuto-check passed
  • Turns a natural-language description of routing intent into a valid Lemonade collection.router policy JSON.

    408 GitHub stars~4k tokensUpdated yesterday
    AI & LLM EngineeringAuto-check passed

More from gtmagents/gtm-agents

All 121 skills in this repo
  • Enrichment Data Sourcing

    gtmagents/gtm-agents

    Chooses and orders data providers for email, phone, company and intent enrichment, with waterfall sequences and credit-saving tactics across 150+ sources.

    414 GitHub starsUsed in 2 repos~2.4k tokens
    Auto-check passed
  • Cold Email Personalization

    gtmagents/gtm-agents

    Writes personalized B2B cold emails from prospect research, using a gated workflow, a quality rubric and follow-up sequences.

    414 GitHub starsUsed in 1 repo~853 tokens
    Auto-check passed
  • Discovery Call Playbook

    gtmagents/gtm-agents

    Structures sales discovery calls with a PREP routine, a five-part call flow, a question bank, a qualification scorecard and a follow-up recap email.

    414 GitHub starsUsed in 1 repo~368 tokens
    Auto-check passed
  • Frames marketing automation around lifecycle stages, signals, touches and SLAs, with worksheets for onboarding, expansion, renewal and churn-prevention journeys.

    414 GitHub starsUsed in 1 repo~927 tokens
    Auto-check passed
  • Social Selling

    gtmagents/gtm-agents

    A skill your agent uses when engaging prospects through LinkedIn, communities, and social channels to spark warm conversations and meetings.

    414 GitHub starsUsed in 1 repo~431 tokens
    Auto-check passed
  • Drip Campaigns

    gtmagents/gtm-agents

    A skill your agent uses when you need to map sequenced nurture flows with pacing, storytelling arcs, and value ladders.

    414 GitHub starsUsed in 1 repo~310 tokens
    Auto-check passed

Questions about Experiment Design Kit

What does Experiment Design Kit do?

Toolkit for structuring hypotheses, variants, guardrails, and measurement plans. Experiment Design Kit is an agent skill from gtmagents/gtm-agents. Toolkit for structuring hypotheses, variants, guardrails, and measurement plans.

When should I use Experiment Design Kit?

Experiment Design Kit fits situations like: tasks that involve Experimental design; tasks that involve LLM guardrails; tasks that involve Go-to-market strategy.

How do I install Experiment Design Kit in Claude Code?

Run `npx skills add gtmagents/gtm-agents --skill experiment-design-kit -a claude-code`. Or copy the skill folder (plugins/growth-experiments/skills/experiment-design-kit in gtmagents/gtm-agents) into .claude/skills/experiment-design-kit in your project. Claude Code loads it when a task matches its description.

How do I install Experiment Design Kit in Codex?

Run `npx skills add gtmagents/gtm-agents --skill experiment-design-kit -a codex`. Or copy the skill folder (plugins/growth-experiments/skills/experiment-design-kit in gtmagents/gtm-agents) into .agents/skills/experiment-design-kit in your project. Codex loads it when a task matches its description.

Can I use Experiment Design Kit in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add gtmagents/gtm-agents --skill experiment-design-kit -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/experiment-design-kit, .gemini/skills/experiment-design-kit, .github/skills/experiment-design-kit and .opencode/skills/experiment-design-kit in your project.

What does Experiment Design Kit need to run?

SKILL.md names no scripts, command-line tools or credentials: Experiment Design Kit is instructions for the agent only.

Does Experiment Design Kit access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is Experiment Design Kit safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Experiment Design Kit use?

Experiment Design Kit is published under the Apache-2.0 licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Experiment Design Kit use?

About 837 tokens (SKILL.md is roughly 3.3k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Experiment Design Kit?

Skills that share tags, products or a category with Experiment Design Kit: Data Scientist (nagisanzenin/production-grade, 181 stars), Phi Prompt Guard (aipoch/medical-research-skills, 1.9k stars), Acm Conference On Fairness Accountability And Transparency (franklee16/academic-research-skills, 223 stars) and Aisafetyhot (wuyoscar/AISafetyHot-Hub, 827 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Experiment Design Kit?

gtmagents (a GitHub user) maintains it in gtmagents/gtm-agents, which has 414 GitHub stars. The repository holds 121 skills in this directory. The repository was last updated on April 3, 2026.

Source: gtmagents/gtm-agents on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.