Agent skill

Define Hypothesis

by product-on-purpose in product-on-purpose/pm-skills

Defines a testable hypothesis with clear success metrics and a validation approach.

Apache-2.0Auto-check passedMarketing & SEO

Install Define Hypothesis

skills CLI
$ npx skills add product-on-purpose/pm-skills --skill define-hypothesis -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install product-on-purpose/pm-skills define-hypothesis --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/product-on-purpose/pm-skills.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/define-hypothesis .claude/skills/define-hypothesis && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
define-hypothesis
GitHub stars
715
Token cost
~966 tokens
SKILL.md length
438 words
Files
5 (incl. references)
Skills in repo
68
Repo updated
First seen
Licence
Apache-2.0

At a glance

Defines a testable hypothesis with clear success metrics and a validation approach.

  • Works in 6 steps: State the Belief → Identify the Target User → Define the Expected Outcome → …
  • Forming assumptions to test
  • SKILL.md covers When to Use, When NOT to Use, Instructions and Output Format, plus 2 more sections
  • Instructions only: no scripts, shell commands, URLs or credentials in SKILL.md

What it does

Define Hypothesis is an agent skill from product-on-purpose/pm-skills. Defines a testable hypothesis with clear success metrics and a validation approach. Use when forming assumptions to test or aligning a team on what success looks like, before any experiment is designed. To design the A/B test or experiment that will validate the hypothesis, use measure-experiment-design.

Its SKILL.md is about 970 tokens, which your agent loads only when the skill is triggered. The skill folder holds 6 other files, including reference files (for example `HISTORY.md`, `evals/trigger-fixtures.json` and `references/EXAMPLE.md`).

It sits in Marketing & SEO, covering A/B testing, Experimental design and Product metrics. The repository describes itself as: 68 plug-and-play, best-practice product management skills for AI agents: 30 Triple Diamond phase + 11 foundation + 12 utility + 15 tool (Foundation Sprint + Design Sprint). Plus… The licence is Apache-2.0.

When your agent uses it

  • Forming assumptions to test
  • Aligning a team on what success looks like
  • Before any experiment is designed

Example prompts

  • “Use the define-hypothesis skill to define a testable hypothesis with clear success metrics and a validation approach”
  • “/define-hypothesis”

Workflow steps

6 steps, taken from the first numbered list in SKILL.md.

  1. State the Belief
  2. Identify the Target User
  3. Define the Expected Outcome
  4. Set Success Metrics
  5. Describe Validation Approach
  6. Document Risks and Assumptions

What it can do on your machine

Read from SKILL.md and the folder at commit 1cef1a9. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    No scripts in the folder and no shell commands in SKILL.md.

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Define Hypothesis loads about 966 tokens when it runs, and up to ~3k if it reads all its reference files. Until then it costs about 81 tokens; SKILL.md has 438 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~81
When it runs · the whole SKILL.md, loaded when a task matches
~966
With references · SKILL.md plus every file in references/, read only if the agent opens them
~3k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from product-on-purpose/pm-skills at commit 1cef1a9, republished under its Apache-2.0 licence (© product-on-purpose). 438 words, ~966 tokens.

Download SKILL.mdSave it as .claude/skills/define-hypothesis/SKILL.md (or your agent's skills folder). This skill also uses 4 other files; get the full folder from GitHub.
name
define-hypothesis
description
Defines a testable hypothesis with clear success metrics and a validation approach. Use when forming assumptions to test or aligning a team on what success looks like, before any experiment is designed. To design the A/B test or experiment that will validate the hypothesis, use measure-experiment-design.
license
Apache-2.0
metadata.phase
define
metadata.version
2.1.0
metadata.updated
2026-06-10
metadata.category
ideation
metadata.frameworks
triple-diamond, lean-startup, design-thinking
metadata.author
product-on-purpose
<!-- PM-Skills | https://github.com/product-on-purpose/pm-skills | Apache 2.0 -->

Hypothesis

A hypothesis is a testable prediction about how a change will affect user behavior or business outcomes. It transforms assumptions into explicit statements that can be validated or invalidated through experimentation. Well-formed hypotheses prevent teams from building features based on untested beliefs and create shared understanding of what success looks like.

When to Use

  • After problem framing, before committing to a solution
  • When designing experiments or A/B tests
  • When team members have differing assumptions about user behavior
  • Before investing significant engineering resources in a feature
  • When pivoting direction and need to validate the new approach

When NOT to Use

  • You are ready to design the actual A/B test (variants, sample size, duration) -> use measure-experiment-design; this skill frames what to test, not how
  • The problem itself is still unframed -> use define-problem-statement first
  • You want to organize many assumptions and ideas into a discovery structure -> use define-opportunity-tree
  • The team needs the full business-model picture, not one testable claim -> use foundation-lean-canvas

Instructions

When asked to create a hypothesis, follow these steps:

  1. State the Belief Articulate what you believe will happen. Use the structured format: "We believe that [action/change] for [target user] will [expected outcome]." Be specific about the intervention - vague hypotheses can't be tested.

  2. Identify the Target User Define who this hypothesis applies to. A hypothesis about "users" is too broad. Specify the segment: new users in their first week, power users with 10+ sessions, churned users returning, etc.

  3. Define the Expected Outcome What behavior change or result do you expect? Frame it in terms of user actions (complete onboarding, make a purchase, return within 7 days) rather than internal metrics when possible.

  4. Set Success Metrics Choose a primary metric that directly measures the expected outcome. Include secondary metrics that provide context and guardrail metrics that ensure you're not causing harm elsewhere.

  5. Describe Validation Approach How will you test this hypothesis? A/B test, user interviews, prototype testing, cohort analysis? Be specific about sample size, duration, and statistical requirements.

  6. Document Risks and Assumptions What could invalidate this hypothesis beyond the test results? What are you assuming to be true that you haven't validated?

Show full SKILL.md (82 more words)Show less

Output Format

Use the template in references/TEMPLATE.md to structure the output. A complete hypothesis document fills every template section: Hypothesis Statement; Background & Rationale; Target User Segment; Success Metrics; Validation Approach; Risks & Assumptions; and Timeline.

Quality Checklist

Before finalizing, verify:

  • Hypothesis is falsifiable (possible to prove wrong)
  • Success metric has a specific numeric target
  • Target user segment is clearly defined
  • Validation approach is practical and time-bound
  • Pass/fail criteria are unambiguous
  • Hypothesis doesn't assume the solution works

Examples

See references/EXAMPLE.md for a completed example.

© product-on-purpose, Apache-2.0. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

SKILL.md and 4 other files (references) in skills/define-hypothesis of product-on-purpose/pm-skills.

  • SKILL.md
  • HISTORY.md
  • evals/trigger-fixtures.json
  • references/EXAMPLE.md
  • references/TEMPLATE.md

Open the folder on GitHubat commit 1cef1a9

Compare with similar skills

Define Hypothesis next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Define Hypothesis compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Define Hypothesis this skillproduct-on-purpose/pm-skills715—~966Automated safety check: PassApache-2.0
Ab Testingericrisco/rsc-harness174—~2.4kAutomated safety check: PassMIT
Craft Experiment Designamplitude/builder-skills160—~522Automated safety check: PassNone
Ab Test Analyzeririnabuht12-oss/marketing-skills4k—~1.4kAutomated safety check: PassNone
A B Test DesignOwl-Listener/designer-skills2.9k1 repos~472Automated safety check: PassMIT
Send Experiment Designeraaron-he-zhu/aaron-marketing-skills2.9k2 repos~4.1kAutomated safety check: PassApache-2.0

Similar skills

  • Ab Testing

    ericrisco/rsc-harness

    A skill your agent uses when designing or analyzing a controlled experiment — falsifiable hypothesis, sample size from an MDE, reading significance/CI/power, CUPED, or rescuing tests that won't go…

    174 GitHub stars~2.4k tokensUpdated 2 days ago
    Marketing & SEOAuto-check passed
  • Craft Experiment Design

    amplitude/builder-skills

    Write a hypothesis, define success metrics, and plan a holdout strategy.

    160 GitHub stars~522 tokensUpdated 2 mo ago
    Research & ScienceAuto-check passed
  • Ab Test Analyzer

    irinabuht12-oss/marketing-skills

    Statistical significance calculator for A/B test results with sample size requirements, segment breakdowns, and hypothesis generation.

    4k GitHub stars~1.4k tokensUpdated 16 days ago
    Marketing & SEOAuto-check passed
  • A B Test Design

    Owl-Listener/designer-skills

    Design an A/B experiment — hypothesis, variants, primary metric, and sample size.

    2.9k GitHub starsUsed in 1 repo~472 tokens
    Marketing & SEOAuto-check passed
  • Send Experiment Designer

    aaron-he-zhu/aaron-marketing-skills

    A skill your agent uses when the user asks to "design an email A/B test", "set up a multivariate subject/CTA test", "run a send-time test", "build a hold-out group", or "is this email result…

    2.9k GitHub starsUsed in 2 repos~4.1k tokens
    Marketing & SEOAuto-check passed
  • Official

    Content experimentation and A/B testing guidance covering experiment design, hypotheses, metrics, sample size, statistical foundations, CMS-managed variants, and common analysis pitfalls.

    188 GitHub starsUsed in 1 repo~469 tokens
    Marketing & SEOAuto-check passed

More from product-on-purpose/pm-skills

All 68 skills in this repo
  • Define Jtbd Canvas

    product-on-purpose/pm-skills

    Creates a Jobs to be Done canvas capturing the functional, emotional, and social dimensions of a customer job.

    715 GitHub stars~1.1k tokensUpdated yesterday
    Auto-check passed
  • Define Opportunity Tree

    product-on-purpose/pm-skills

    Creates an opportunity solution tree connecting a desired outcome to customer opportunities and candidate solutions, preventing solution-first jumps in continuous discovery.

    715 GitHub stars~1.1k tokensUpdated yesterday
    Auto-check passed
  • Define Problem Statement

    product-on-purpose/pm-skills

    Creates a clear problem framing document with user impact, business context, and success criteria.

    715 GitHub stars~932 tokensUpdated yesterday
    Auto-check passed
  • Deliver Acceptance Criteria

    product-on-purpose/pm-skills

    Generates structured Given/When/Then acceptance criteria for a user story or feature slice, covering the happy path, key failure scenarios, and non-functional expectations in testable form.

    715 GitHub stars~1k tokensUpdated yesterday
    Auto-check passed
  • Deliver Launch Checklist

    product-on-purpose/pm-skills

    Creates a cross-functional pre-launch checklist covering engineering, design, marketing, support, legal, and operations readiness, with owners, dates, and go/no-go criteria so nothing is missed…

    715 GitHub stars~970 tokensUpdated yesterday
    Auto-check passed
  • Deliver Prd

    product-on-purpose/pm-skills

    Creates a comprehensive Product Requirements Document that aligns stakeholders on what to build, why, and how success will be measured.

    715 GitHub stars~2k tokensUpdated yesterday
    Auto-check passed

Questions about Define Hypothesis

What does Define Hypothesis do?

Defines a testable hypothesis with clear success metrics and a validation approach. Define Hypothesis is an agent skill from product-on-purpose/pm-skills. Defines a testable hypothesis with clear success metrics and a validation approach.

When should I use Define Hypothesis?

Define Hypothesis fits situations like: forming assumptions to test; aligning a team on what success looks like; before any experiment is designed.

How do I install Define Hypothesis in Claude Code?

Run `npx skills add product-on-purpose/pm-skills --skill define-hypothesis -a claude-code`. Or copy the skill folder (skills/define-hypothesis in product-on-purpose/pm-skills) into .claude/skills/define-hypothesis in your project. Claude Code loads it when a task matches its description.

How do I install Define Hypothesis in Codex?

Run `npx skills add product-on-purpose/pm-skills --skill define-hypothesis -a codex`. Or copy the skill folder (skills/define-hypothesis in product-on-purpose/pm-skills) into .agents/skills/define-hypothesis in your project. Codex loads it when a task matches its description.

Can I use Define Hypothesis in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add product-on-purpose/pm-skills --skill define-hypothesis -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/define-hypothesis, .gemini/skills/define-hypothesis, .github/skills/define-hypothesis and .opencode/skills/define-hypothesis in your project.

What does Define Hypothesis need to run?

SKILL.md names no scripts, command-line tools or credentials: Define Hypothesis is instructions for the agent only.

Does Define Hypothesis access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is Define Hypothesis safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Define Hypothesis use?

Define Hypothesis is published under the Apache-2.0 licence (declared in SKILL.md). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Define Hypothesis use?

About 966 tokens (SKILL.md is roughly 3.9k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full. Its references folder adds about 2.1k tokens, read only when the agent opens those files.

What are the alternatives to Define Hypothesis?

Skills that share tags, products or a category with Define Hypothesis: Ab Testing (ericrisco/rsc-harness, 174 stars), Craft Experiment Design (amplitude/builder-skills, 160 stars), Ab Test Analyzer (irinabuht12-oss/marketing-skills, 4k stars) and A B Test Design (Owl-Listener/designer-skills, 2.9k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Define Hypothesis?

product-on-purpose (a GitHub organization) maintains it in product-on-purpose/pm-skills, which has 715 GitHub stars. The repository holds 68 skills in this directory. The repository was last updated on October 8, 2026.

Source: product-on-purpose/pm-skills on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.