Agent skill

Experimentation And Ab Testing

by social-media-skills in social-media-skills/skills

A/B testing and experimentation for social media content — evidence over opinion, at honest organic-scale rigor.

MITAuto-check passedMarketing & SEO

Install Experimentation And Ab Testing

skills CLI
$ npx skills add social-media-skills/skills --skill experimentation-and-ab-testing -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install social-media-skills/skills experimentation-and-ab-testing --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/social-media-skills/skills.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/experimentation-and-ab-testing .claude/skills/experimentation-and-ab-testing && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
experimentation-and-ab-testing
GitHub stars
134
Token cost
~1.5k tokens
SKILL.md length
545 words
Files
6 (incl. references)
Skills in repo
106
Repo updated
First seen
Licence
MIT

At a glance

A/B testing and experimentation for social media content — evidence over opinion, at honest organic-scale rigor.

  • Works in 2 steps: brand-profile — voice/format constraints… → goals-and-kpis — the KPI/primary metric…
  • Someone wants to A/B test
  • SKILL.md covers The POV: evidence, not vibes, Read these first, The framework: TEST and What to test (highest…, plus 4 more sections
  • Instructions only: no scripts, shell commands, URLs or credentials in SKILL.md

What it does

Experimentation And Ab Testing is an agent skill from social-media-skills/skills. A/B testing and experimentation for social media content — evidence over opinion, at honest organic-scale rigor. Use when someone wants to "A/B test" or "split test" content, "test which version/hook/time/caption/thumbnail works," "set up an experiment," "what should I test," or to settle a content debate with evidence instead of opinion. Designs disciplined organic tests: one variable, controlled context, a decision rule set BEFORE publishing, enough duration and repetitions to separate signal from noise. Uses…

Its SKILL.md is about 1.5k tokens, which your agent loads only when the skill is triggered. The skill folder holds 7 other files, including reference files (for example `evals/evals.json`, `references/experimentation-2026-reality.md` and `references/scope-and-connections.md`).

It sits in Marketing & SEO, covering A/B testing. It works with YouTube. The repository describes itself as: 106 social media skills for AI agents - strategy, writing, video, design, platform growth, publishing, and analytics. Works with Claude, Cursor, OpenClaw, Hermes & 40+ agents. The licence is MIT.

When your agent uses it

  • Someone wants to A/B test
  • Split test content
  • Test which version/hook/time/caption/thumbnail works
  • Set up an experiment

Example prompts

  • “A/B test”
  • “split test”
  • “test which version/hook/time/caption/thumbnail works,”
  • “/experimentation-and-ab-testing”

Workflow steps

2 steps, taken from the first numbered list in SKILL.md.

  1. brand-profile — voice/format constraints for the variants.
  2. goals-and-kpis — the KPI/primary metric the test must move.

What it can do on your machine

Read from SKILL.md and the folder at commit 6e30eeb. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    No scripts in the folder and no shell commands in SKILL.md.

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Experimentation And Ab Testing loads about 1.5k tokens when it runs, and up to ~4.5k if it reads all its reference files. Until then it costs about 249 tokens; SKILL.md has 545 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~249
When it runs · the whole SKILL.md, loaded when a task matches
~1.5k
With references · SKILL.md plus every file in references/, read only if the agent opens them
~4.5k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from social-media-skills/skills at commit 6e30eeb, republished under its MIT licence (© social-media-skills). 545 words, ~1,456 tokens.

Download SKILL.mdSave it as .claude/skills/experimentation-and-ab-testing/SKILL.md (or your agent's skills folder). This skill also uses 5 other files; get the full folder from GitHub.
name
experimentation-and-ab-testing
description
A/B testing and experimentation for social media content — evidence over opinion, at honest organic-scale rigor. Use when someone wants to "A/B test" or "split test" content, "test which version/hook/time/caption/thumbnail works," "set up an experiment," "what should I test," or to settle a content debate with evidence instead of opinion. Designs disciplined organic tests: one variable, controlled context, a decision rule set BEFORE publishing, enough duration and repetitions to separate signal from noise. Uses the TEST framework. Reads brand-profile + goals-and-kpis first. It DESIGNS the test and drafts variants; WoopSocial schedules them as controlled sequential posts (exception: YouTube's native Test & Compare); the result is read from native analytics via analytics-and-reporting. Organic can't reach true statistical significance; nothing is fabricated or p-hacked. Distinct from analytics-and-reporting (measures) and goals-and-kpis (sets targets).
version
1.0.0

experimentation-and-ab-testing

The causation engine — manipulate one variable under controlled conditions to learn what actually moves a KPI. This skill designs the test and drafts variants; scheduling-and-queue → WoopSocial publishes them; analytics-and-reporting reads the result.

The POV: evidence, not vibes

Most "testing" on social is vibes — post two things, eyeball the likes, declare a winner, learn nothing. Real experimentation turns guesses into evidence: change one variable, control everything else, set the decision rule before you publish, and run it long and often enough to separate signal from noise. Organic can't give clean statistical significance (small samples, an algorithm in the middle), so you compensate with tighter controls, a ~20%+ effect threshold, guardrail metrics, and 3–5 repetitions — and treat a single viral post as noise, not a strategy.

Read these first

  1. brand-profile — voice/format constraints for the variants.
  2. goals-and-kpis — the KPI/primary metric the test must move.

The framework: TEST

(Depth: references/the-test-framework.md.)

  • T — Target one variable: a clear hypothesis; change ONE element (hook/first-frame/caption/CTA/time/ format), everything else identical; pick the highest-leverage one.
  • E — Establish the decision rule first: set the primary metric + win threshold + guardrail before publishing ("B wins if reach +15% and saves/reach not worse"); no post-hoc rationalizing.
  • S — Set controls + sample: same platform/format/topic/length/window; run ≥7 days (small accounts 2–4 weeks); judge on a ~20%+ consistent effect (a tie = "test elsewhere").
  • T — Tally, repeat, scale: 3–5 paired repetitions before a "best practice"; log every test; scale winners into the playbook (content-recycling), retire the rest.

What to test (highest leverage, in your control)

Hook/first-frame (short video) → posting time (easy) → format → caption/CTA → thumbnail → hashtags — always tied to the KPI; test what's in your control, not algorithm-dependent factors. Run a 30-day sprint with one test always running. Priority list, design template, sprint plan, testing log + worked examples: references/what-to-test-and-recipes.md. Full method + rules: references/experimentation-2026-reality.md.

Show full SKILL.md (249 more words)Show less

Honest scope (never violate)

  • Organic isn't lab-grade — results are directional; compensate with controls + effect-size + repetition, not p-value theater.
  • WoopSocial has no A/B/audience-split surface → organic testing = controlled sequential posts; the agent designs + drafts variants + schedules; the primary metric is read from native analytics (analytics-and-reporting). One true native split exists: YouTube's Test & Compare (YouTube Studio, long-form, not Shorts) — up to 3 titles, thumbnails, or title+thumbnail combos; use it for YouTube title/thumbnail tests instead of sequential posts. (verify-quarterly)
  • No p-hacking / HARKing / cherry-picking — decision rule pre-set; a multi-variable change can't be pinned on one element; one post/one day is noise. Never fabricate a result; a tie is valid. (Scope, the loop role + connections: references/scope-and-connections.md.)

Distinct from its siblings (route correctly)

experimentation (this) = manipulate one variable to establish causation · analytics-and-reporting = observe/measure what happened · goals-and-kpis = set the target/primary metric · content-recycling = scale proven winners · viral-reverse-engineering = explain a past post (hindsight) vs testing forward.

Where this connects

Reads first: brand-profile, goals-and-kpis. Variants drafted via: hook-writer, caption-writer, reels-script/tiktok-script, carousel-writer, image-prompt/ideogram/nano-banana, thumbnail-design. Readout: analytics-and-reporting (native analytics). Scale/plan: content-recycling, social-strategy, content-calendar/batch-content-plan, every *-growth skill. Publish variants: scheduling-and-queue → WoopSocial (controlled sequential posts).

Definition of done

A clear hypothesis testing ONE variable tied to a KPI; identical controlled context; a primary metric + win threshold + guardrail set before publishing; duration ≥7 days (2–4 weeks small accounts) and 3–5 paired repetitions; results read from native analytics and judged on a ~20%+ consistent effect (ties acknowledged); winners logged and scaled to content-recycling/strategy; organic limits stated, nothing fabricated or p-hacked, correctly distinguished from analytics-and-reporting and goals-and-kpis.

© social-media-skills, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

SKILL.md and 5 other files (references) in skills/experimentation-and-ab-testing of social-media-skills/skills.

  • SKILL.md
  • evals/evals.json
  • references/experimentation-2026-reality.md
  • references/scope-and-connections.md
  • references/the-test-framework.md
  • references/what-to-test-and-recipes.md

Open the folder on GitHubat commit 6e30eeb

Compare with similar skills

Experimentation And Ab Testing next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Experimentation And Ab Testing compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Experimentation And Ab Testing this skillsocial-media-skills/skills134—~1.5kAutomated safety check: PassMIT
Video Title Optimizernicepkg/ai-workflow285—~1.8kAutomated safety check: PassMIT
Title CraftTheCraigHewitt/skills159—~2.1kAutomated safety check: PassMIT
Youtube Thumbnailsericrisco/rsc-harness180—~2.4kAutomated safety check: PassMIT
Video Hook Generatornicepkg/ai-workflow285—~2kAutomated safety check: PassMIT
Ab Testingcoreyhaines31/marketingskills54k3 repos~3.1kAutomated safety check: PassMIT

Similar skills

  • Video Title Optimizer

    nicepkg/ai-workflow

    Optimize video titles for maximum click-through rate (CTR) and YouTube/TikTok SEO.

    285 GitHub stars~1.8k tokensUpdated 8 mo ago
    Marketing & SEOAuto-check passed
  • Title Craft

    TheCraigHewitt/skills

    When the user wants to write YouTube video titles, optimize existing titles, A/B test title variants, or improve click-through rate.

    159 GitHub stars~2.1k tokensUpdated 4 mo ago
    Media & CreativeAuto-check passed
  • Youtube Thumbnails

    ericrisco/rsc-harness

    A skill your agent uses when designing or fixing a YouTube thumbnail, making variants for an A/B test, or working out which thumbnail style wins on a channel — including the trap where CTR looks…

    180 GitHub stars~2.4k tokensUpdated yesterday
    Media & CreativeAuto-check passed
  • Video Hook Generator

    nicepkg/ai-workflow

    Generate attention-grabbing hooks for the first 3 seconds of videos.

    285 GitHub stars~2k tokensUpdated 8 mo ago
    Media & CreativeAuto-check passed
  • Ab Testing

    coreyhaines31/marketingskills

    When the user wants to plan, design, or implement an A/B test or experiment, or build a growth experimentation program.

    54k GitHub starsUsed in 3 repos~3.1k tokens
    Marketing & SEOAuto-check passed
  • Google SEO APIs

    AgriciDaniel/claude-seo

    Pulls real Google data for SEO work: Search Console, PageSpeed Insights, CrUX field data, the Indexing API and GA4 organic traffic, through /seo google commands.

    19k GitHub starsUsed in 1 repo~4.2k tokens
    Marketing & SEOAuto-check passed

More from social-media-skills/skills

All 106 skills in this repo
  • AI Image Editing

    social-media-skills/skills

    The AI image-editing router — inpainting/object removal, background removal, upscaling, outpainting, old-photo restoration, and retouch, routed task-first to the right engine.

    134 GitHub stars~2.1k tokensUpdated 9 days ago
    Auto-check passed
  • AI Music And Sound

    social-media-skills/skills

    The AI music + sound-design skill for social -- original/licensed audio beds and sound design for Reels/TikToks/Shorts/videos.

    134 GitHub stars~1.9k tokensUpdated 9 days ago
    Auto-check passed
  • AI Search Optimization

    social-media-skills/skills

    A skill your agent uses to get a brand and its content CITED and RECOMMENDED by AI answer engines — the GEO (Generative Engine Optimization) / AI-search-visibility skill.

    134 GitHub stars~2k tokensUpdated 9 days ago
    Auto-check passed
  • AI Video

    social-media-skills/skills

    The model-agnostic AI-video router and brief — the counterpart to image-prompt.

    134 GitHub stars~1.3k tokensUpdated 9 days ago
    Auto-check passed
  • AI Voiceover

    social-media-skills/skills

    The AI narration / voiceover mini-skill (ElevenLabs-led). An agent skill from social-media-skills/skills.

    134 GitHub stars~1.1k tokensUpdated 9 days ago
    Auto-check passed
  • Analytics And Reporting

    social-media-skills/skills

    Social media analytics and reporting — read native platform data honestly and turn it into next actions.

    134 GitHub stars~1.4k tokensUpdated 9 days ago
    Auto-check passed

Works with

Categories

Questions about Experimentation And Ab Testing

What does Experimentation And Ab Testing do?

A/B testing and experimentation for social media content — evidence over opinion, at honest organic-scale rigor. Experimentation And Ab Testing is an agent skill from social-media-skills/skills. A/B testing and experimentation for social media content — evidence over opinion, at honest organic-scale rigor.

When should I use Experimentation And Ab Testing?

Experimentation And Ab Testing fits situations like: someone wants to A/B test; split test content; test which version/hook/time/caption/thumbnail works; set up an experiment.

How do I install Experimentation And Ab Testing in Claude Code?

Run `npx skills add social-media-skills/skills --skill experimentation-and-ab-testing -a claude-code`. Or copy the skill folder (skills/experimentation-and-ab-testing in social-media-skills/skills) into .claude/skills/experimentation-and-ab-testing in your project. Claude Code loads it when a task matches its description.

How do I install Experimentation And Ab Testing in Codex?

Run `npx skills add social-media-skills/skills --skill experimentation-and-ab-testing -a codex`. Or copy the skill folder (skills/experimentation-and-ab-testing in social-media-skills/skills) into .agents/skills/experimentation-and-ab-testing in your project. Codex loads it when a task matches its description.

Can I use Experimentation And Ab Testing in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add social-media-skills/skills --skill experimentation-and-ab-testing -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/experimentation-and-ab-testing, .gemini/skills/experimentation-and-ab-testing, .github/skills/experimentation-and-ab-testing and .opencode/skills/experimentation-and-ab-testing in your project.

What does Experimentation And Ab Testing need to run?

SKILL.md names no scripts, command-line tools or credentials: Experimentation And Ab Testing is instructions for the agent only.

Does Experimentation And Ab Testing access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is Experimentation And Ab Testing safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Experimentation And Ab Testing use?

Experimentation And Ab Testing is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Experimentation And Ab Testing use?

About 1.5k tokens (SKILL.md is roughly 5.8k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full. Its references folder adds about 3k tokens, read only when the agent opens those files.

What are the alternatives to Experimentation And Ab Testing?

Skills that share tags, products or a category with Experimentation And Ab Testing: Video Title Optimizer (nicepkg/ai-workflow, 285 stars), Title Craft (TheCraigHewitt/skills, 159 stars), Youtube Thumbnails (ericrisco/rsc-harness, 180 stars) and Video Hook Generator (nicepkg/ai-workflow, 285 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Experimentation And Ab Testing?

social-media-skills (a GitHub organization) maintains it in social-media-skills/skills, which has 134 GitHub stars. The repository holds 106 skills in this directory. The repository was last updated on October 1, 2026.

Source: social-media-skills/skills on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.