Topic · Marketing & SEO
Best A/B testing skills for Claude Code, Codex and other agents.
- skills
- 267
- official
- 9
A/B testing skills, ranked
Ranked by score. Sort bymost stars,trending,newest,recently updated
| # | Skill | Repository | Stars | Used in | Tokens | Auto-check | Licence | Updated |
|---|---|---|---|---|---|---|---|---|
| 1 | When the user wants to set up, improve, or audit analytics tracking and measurement. | Nexus-JPF/ | 869 | 6 repos | ~2.2k | Automated safety check: Pass | MIT | 4 days ago |
| 2 | Pre-pipeline aggregator that scans AI agent cache directories (.claude, .cursor, .antigravity, .openclaw) or any user-specified directory for experimentation logs, extracts insights and numeric… | Ar9av/ | 676 | 2 repos | ~3.5k | Automated safety check: Pass | Unknown | 16 days ago |
| 3 | When the user wants to plan, design, or implement an A/B test or experiment. | freekmurze/ | 1k | 15 repos | ~1.8k | Automated safety check: Pass | No licence | 2 days ago |
| 4 | Rigor Improve / Rigor Explore run leaf skill for bounded exploratory evidence in deep learning research repositories. | lllllllama/ | 497 | 2 repos | ~833 | Automated safety check: Pass | MIT | 14 days ago |
| 5 | A skill your agent uses when the user asks to "design an A/B test", "set up a creative/landing test", "run an incrementality test", or "is this result statistically and practically material?"… | aaron-he-zhu/ | 2.9k | 2 repos | ~2.8k | Automated safety check: Pass | Apache-2.0 | today |
| 6 | When the user wants to plan, design, or implement an A/B test or experiment, or build a growth experimentation program. | Cesarjoquin/ | 199 | 2 repos | ~2.8k | Automated safety check: Pass | MIT | 23 days ago |
| 7 | World-class data science skill for statistical modeling, experimentation, causal inference, and advanced analytics. | Raidriar7170/ | 125 | 6 repos | ~1.4k | Automated safety check: Pass | MIT | 11 days ago |
| 8 | Writes and improves title tags, meta descriptions, Open Graph and Twitter card tags for click-through, with character counts and A/B test variants. | nowork-studio/ | 3.9k | 1 repo | ~2.7k | Automated safety check: Pass | MIT | 6 days ago |
| 9 | Statistical significance calculator for A/B test results with sample size requirements, segment breakdowns, and hypothesis generation. | irinabuht12-oss/ | 3.8k | — | ~1.4k | Automated safety check: Pass | No licence | 14 days ago |
| 10 | Generate high-CTR YouTube thumbnails using Nano Banana 2 via the Arcads external API. | krusemediallc/ | 1.6k | — | ~2.4k | Automated safety check: Notes | MIT | 15 days ago |
| 11 | When the user wants to A/B test App Store product page elements to improve conversion rate. | appeeky/ | 2.1k | — | ~1.8k | Automated safety check: Pass | MIT | yesterday |
| 12 | Benchmark and profile the Dynamo frontend (dynamo.frontend HTTP + tokenizer + KV router) against mock workers (dynamo.mocker). | ai-dynamo/ | 8.2k | — | ~3.5k | Automated safety check: Notes | Apache-2.0 | today |
| 13 | Translate an existing Remotion (React-based) video composition into a HyperFrames HTML composition. | boraoztunc/ | 393 | — | ~2.2k | Automated safety check: Pass | Apache-2.0 | 1 mo ago |
| 14 | Hack, modify, and translate ColecoVision and Super Game Module ROMs using the Gearcoleco emulator MCP server. | drhelius/ | 141 | — | ~3.9k | Automated safety check: Pass | GPL-3.0 | yesterday |
| 15 | Transform Claude Code into an AI Scientist that orchestrates research workflows using tree-based hypothesis exploration. | sundial-org/ | 152 | — | ~2.5k | Automated safety check: Pass | No licence | 2 mo ago |
| 16 | Defines a testable hypothesis with clear success metrics and a validation approach. | product-on-purpose/ | 713 | — | ~966 | Automated safety check: Pass | Apache-2.0 | 3 days ago |
| 17 | A skill your agent uses when the user has changed a prompt (system prompt, RAG template, agent instruction, etc.) and wants to know whether the candidate is better or worse than the baseline. | agentscope-ai/ | 867 | — | ~2.8k | Automated safety check: Pass | Apache-2.0 | 26 days ago |
| 18 | 18.Skill Forge Ultimate Claude Code skill creator and architect. An agent skill from AgriciDaniel/skill-forge. | AgriciDaniel/ | 177 | — | ~1.9k | Automated safety check: Notes | MIT | 6 mo ago |
| 19 | When the user wants to set up, improve, or audit analytics tracking and measurement. | freekmurze/ | 1k | 12 repos | ~2k | Automated safety check: Pass | No licence | 2 days ago |
| 20 | When the user wants to set up, interpret, or improve their app analytics and tracking. | appeeky/ | 2.1k | — | ~1.6k | Automated safety check: Pass | MIT | yesterday |
| 21 | Rigor Explore compatible skill slug for meaningful and potentially novel deep learning research candidates. | lllllllama/ | 497 | 1 repo | ~1.7k | Automated safety check: Pass | MIT | 14 days ago |
| 22 | 22.Kapso Optimize code using KAPSO (Knowledge-Grounded Optimization). | Leeroo-AI/ | 120 | — | ~642 | Automated safety check: Pass | MIT | yesterday |
| 23 | Validates an experiment's setup, works out lift, p-value and confidence interval from A/B test data, and recommends whether to ship, extend or stop. | phuryn/ | 27k | — | ~893 | Automated safety check: Pass | MIT | 23 days ago |
| 24 | A skill your agent uses when the user asks to "write the email", "draft subject lines", or "build email creative"; produces the pre-click unit — subject-line variants + preheader, body copy, one… | aaron-he-zhu/ | 2.9k | 2 repos | ~3.1k | Automated safety check: Pass | Apache-2.0 | today |
| 25 | Add full builder API support (@tag, @parse, @split) for a TTIR op. | tenstorrent/ | 311 | — | ~1.6k | Automated safety check: Pass | Apache-2.0 | today |
| 26 | Design an A/B experiment — hypothesis, variants, primary metric, and sample size. | Owl-Listener/ | 2.9k | 1 repo | ~472 | Automated safety check: Pass | MIT | 1 mo ago |
| 27 | 27.Ab Test Plan Design a statistically rigorous A/B or multivariate test plan — If/Then/Because hypothesis, control and variant specs, required sample size per variant (absolute vs relative MDE via… | indranilbanerjee/ | 854 | 1 repo | ~2k | Automated safety check: Pass | MIT | 3 days ago |
| 28 | Validate multi-line log stitching behavior for an ama-logs image change. | microsoft/ | 173 | — | ~2.5k | Automated safety check: Pass | Unknown | yesterday |
| 29 | The standard way to run a listening test in HOT-Step - a local HTML score sheet next to the renders where Rob plays each track, scores it 1-5 on named criteria, and the page charts the two score… | scragnog/ | 170 | — | ~1.9k | Automated safety check: Pass | MIT | 2 days ago |
| 30 | PhD-level expertise in data science, statistics, and machine learning. | magnus919/ | 278 | — | ~3.3k | Automated safety check: Pass | MIT | 3 mo ago |
| 31 | When the user wants to design, test, or improve their app icon to increase tap-through rate and conversions in App Store search and browse. | appeeky/ | 2.1k | — | ~1.5k | Automated safety check: Pass | MIT | yesterday |
| 32 | Benchmark Claude Code skill performance with variance analysis, tracking pass rate, execution time, and token usage across iterations. | AgriciDaniel/ | 177 | — | ~1.4k | Automated safety check: Pass | MIT | 6 mo ago |
| 33 | Rigorous A/B test statistical analysis. An agent skill from nimrodfisher/data-analytics-skills. | nimrodfisher/ | 465 | — | ~708 | Automated safety check: Pass | MIT | 12 days ago |
| 34 | Forces heavy internal computation (thinking tokens) before each response. | Lomnus-ai/ | 173 | — | ~9.1k | Automated safety check: Pass | MIT | 5 mo ago |
| 35 | 35.Researcher A skill your agent uses when user wants to optimize or tune something measurable through repeated experiments — "make X faster", "improve [metric]", "find best config", "iterate overnight", "run… | krzysztofdudek/ | 265 | — | ~5.5k | Automated safety check: Warn | MIT | 2 days ago |
| 36 | Craft persuasive, goal-driven emails using AI. An agent skill from ILoveDotNet/ilovedotnet. | ILoveDotNet/ | 155 | — | ~3k | Automated safety check: Pass | CC0-1.0 | yesterday |
| 37 | Interviews you about objective, audience and context, then designs an end-of-article call to action with copy, placement, A/B test plan and accessibility check. | samber/ | 228 | — | ~3.1k | Automated safety check: Pass | MIT | 6 days ago |
| 38 | Make browser automation reliable — anchor every click, type, and scroll to stable DOM selectors and semantic roles instead of screen coordinates, so actions survive layout shifts, A/B tests, and… | breakstageaxe61/ | 125 | — | ~1k | Automated safety check: Pass | MIT | 11 days ago |
| 39 | Run a structured 5-day process to prototype, test, and validate product ideas with real users. | wondelai/ | 2.4k | — | ~3.8k | Automated safety check: Pass | MIT | 27 days ago |
| 40 | 40.Ab Testing A skill your agent uses when designing or analyzing a controlled experiment — falsifiable hypothesis, sample size from an MDE, reading significance/CI/power, CUPED, or rescuing tests that won't go… | ericrisco/ | 156 | — | ~2.4k | Automated safety check: Pass | MIT | yesterday |
| 41 | How to size, write and A/B-test the text of a tool — its description and its inputSchema parameter prose — so that cutting it does not cost call quality, and so that a tool that IS getting called… | DitriXNew/ | 295 | — | ~3k | Automated safety check: Pass | AGPL-3.0 | today |
| 42 | When the user wants to plan, script, produce, or optimize App Store Preview videos or Google Play promo videos — the autoplay videos that show in App Store/Play Store search and product pages. | appeeky/ | 2.1k | — | ~1.6k | Automated safety check: Pass | MIT | yesterday |
| 43 | Edge middleware for EdgeOne Makers — request interception, redirects, rewrites, auth guards, A/B testing, and header injection at the edge (V8 runtime). | TencentEdgeOne/ | 1.9k | 1 repo | ~1.2k | Automated safety check: Pass | MIT | 14 days ago |
| 44 | Calculate A/B test statistical significance. An agent skill from guia-matthieu/clawfu-skills. | guia-matthieu/ | 150 | — | ~1k | Automated safety check: Pass | MIT | 6 days ago |
| 45 | Master comprehensive evaluation strategies for LLM applications, from automated metrics to human evaluation and A/B testing. | davila7/ | 32k | 12 repos | ~3.5k | Automated safety check: Pass | MIT | today |
| 46 | 46.Sap AI Core Guides development with SAP AI Core and SAP AI Launchpad for enterprise AI/ML workloads on SAP BTP. | secondsky/ | 460 | — | ~3.3k | Automated safety check: Pass | GPL-3.0 | 2 days ago |
| 47 | 47.Autoresearch Run Karpathy-style autoresearch optimization on any content. | ericosiu/ | 3.6k | 2 repos | ~2.2k | Automated safety check: Pass | MIT | 15 days ago |
| 48 | Run an explicitly requested Zephyr control/treatment benchmark on the same pre-normalized sample and compare Finelog stage metrics. | marin-community/ | 3.9k | — | ~3.3k | Automated safety check: Pass | Apache-2.0 | today |
Questions, answered from the data.
What is the best A/B testing skill?
Analytics from Nexus-JPF/note-companion ranks first of the 267 A/B testing skills listed here, with the highest score: its repository has 869 GitHub stars, 6 other GitHub owners carry a copy, its SKILL.md loads about 2.2k tokens and it passes the automated safety check with no findings. Next come Agent Research Aggregator and Ab Test Setup.
Which A/B testing skills are official?
9 of the 267 A/B testing skills are official, published by the vendor's own GitHub organization: Multiline Validation, Autoresearch, Content Experimentation Best Practices, Arize Experiment, Flags SDK and 4 more.
How are these skills ranked?
By Skill Navigator score, which combines the GitHub stars of the skill's repository (shared across that repo's skills and discounted for large collections), how many other GitHub owners carry a copy of the skill, and automated SKILL.md quality checks, minus penalties for safety-check warnings and for each further skill from the same repository. Skills that fail the safety check are not listed.
Explore related skills
Category
More topics in Marketing & SEO
- Positioning and messaging620
- Paid advertising433
- AI search optimization383
- Go-to-market strategy375
- Competitor analysis321
- Schema markup289
- Market research275
- On-page SEO246
- Influencer and creator marketing233
- Technical SEO216
- Product launch strategy202
- Keyword research179
- Brand strategy and identity175
- Email marketing169
- Conversion rate optimization168
- Content marketing165
- SEO audit158
- Lead generation150
- Link building142
- Social media marketing124
- Marketing analytics124
- Referral and retention marketing107
- Local SEO91
- Marketing psychology85