Category

Best testing and QA skills for Claude Code, Codex and other agents.

Skills that write, run and debug tests, from TDD and unit tests to browser automation, load tests and QA checklists.
skills
4,904
official
371

Testing & QA skills, ranked

Ranked by score. Sort bymost stars,trending,newest,recently updated

Testing & QA skills, ranked
#SkillRepositoryStarsUsed inTokensAuto-checkLicenceUpdated
1

Browser automation CLI for AI agents. Use when the user needs to interact with websites, including navigating pages, filling forms, clicking…

vercel-labs/agent-browser44k24 repos~864Automated safety check: PassApache-2.0yesterday
2

Explores a web app with the agent-browser CLI to find bugs and UX problems, then writes a report with screenshots, repro videos and step-by-step reproduction for each issue.

vercel-labs/agent-browser44k8 repos~2.7kAutomated safety check: PassApache-2.0yesterday
3

Watches an open GitHub pull request until it merges, handling review comments, diagnosing CI failures and retrying flaky checks along the way.

openinterpreter/openinterpreter69k3 repos~4.2kAutomated safety check: PassApache-2.0today
4

Investigates a session where Superpowers went wrong, reads the transcripts on disk and produces an evidence-cited report, optionally prepared as a bug report for the maintainers.

obra/superpowers296k3 repos~1.7kAutomated safety check: PassMITyesterday
5

Tests local web applications with Python Playwright scripts, checking frontend behavior, capturing screenshots and reading browser console logs.

anthropics/skills180k51 repos~966Automated safety check: PassApache-2.02 days ago
6

Grill the user relentlessly about a plan, decision, or idea.

bestofjs/bestofjs3.1k30 repos~464Automated safety check: PassMIT3 days ago
7

Automates a real browser through short TypeScript scripts that keep page state between runs, for navigating, filling forms, taking screenshots and extracting data.

MemTensor/MemOS12k3 repos~1.7kAutomated safety check: PassApache-2.08 days ago
8
8.Playwright CLIOfficial

Automates browser interactions for web testing, form filling, screenshots, and data extraction.

sanity-io/sanity6.4k18 repos~1.9kAutomated safety check: PassMITtoday
9

Automates Electron desktop apps such as VS Code, Slack or Discord by connecting agent-browser to their Chrome DevTools Protocol port.

vercel-labs/agent-browser44k5 repos~1.7kAutomated safety check: PassApache-2.0yesterday
10

Validates OpenHarness features by running real multi-turn agent loops with live LLM calls against an unfamiliar codebase, checking actual tool execution.

HKUDS/OpenHarness16k1 repo~2.1kAutomated safety check: NotesMIT4 mo ago
11

iOS Simulator control from inside Orca, with the live device view in Orca's emulator pane. Use when driving a booted Apple Simulator on macOS: taps…

stablyai/orca87k1 repo~584Automated safety check: PassApache-2.0today
12

Drives a real browser from the command line with playwright-cli to open pages, interact, mock requests, save state and work with Playwright tests.

github/gh-aw5.3k23 repos~2.8kAutomated safety check: PassMITtoday
13

Investigates failing Pester tests in PowerShell CI jobs by following a six-step workflow from pull request status to documented fix recommendations.

PowerShell/PowerShell56k—~5.1kAutomated safety check: PassMITyesterday
14

Jest patterns for React Native style tests: TDD discipline, mock factory functions, module and GraphQL hook mocking, custom render helpers and anti-patterns to avoid.

ChrisWiles/claude-code-showcase6.1k7 repos~1.5kAutomated safety check: PassNo licence9 mo ago
15

Writes a Playwright end-to-end test from a prompt, runs it against a live Appsmith deployment and retries with fixes up to three times until it passes.

appsmithorg/appsmith41k—~2.9kAutomated safety check: NotesApache-2.0yesterday
16

Diagnosis loop for hard bugs and performance regressions. An agent skill from fossasia/eventyay-interpretation.

fossasia/eventyay-interpretation1.6k31 repos~2.1kAutomated safety check: PassApache-2.02 days ago
17

Enforces red-green-refactor for Rust work, with idiomatic test patterns, a naming convention and a pre-commit gate of cargo fmt, clippy and test.

rtk-ai/rtk83k—~753Automated safety check: NotesApache-2.0yesterday
18

Captures screenshots of the running Mailspring dev app for docs, PRs or visual checks by launching it with a debugging port, driving the UI and clipping to an element.

Foundry376/Mailspring18k—~1.4kAutomated safety check: PassGPL-3.0yesterday
19

Core usage guide for the agent-browser CLI: the snapshot-and-ref workflow for navigating, clicking, filling forms, extracting data and running parallel sessions.

vercel-labs/agent-browser44k4 repos~9.5kAutomated safety check: PassApache-2.0yesterday
20

A skill your agent uses when working inside the ClawTeam repository itself: local development, debugging, reviewing, testing, validating multi-agent flows, or checking whether a code change actually…

HKUDS/ClawTeam5.5k1 repo~1.1kAutomated safety check: PassMIT5 mo ago
21

Walks through an end-to-end smoke test of a DeerFlow deployment: pull the latest code, deploy with Docker or locally, verify services, run health checks and write a report.

bytedance/deer-flow83k—~2.5kAutomated safety check: NotesMITtoday
22

Benchmarks how much CodeGraph helps a coding agent on a real repository, comparing runs with and without it for a chosen local or published version.

colbymchenry/codegraph73k—~950Automated safety check: PassMITtoday
23

Guides changes and reviews of the Cucumber and Playwright end-to-end suite under `e2e/`: feature files, step definitions, support code, tags, locators and assertions.

langgenius/dify158k—~682Automated safety check: PassUnknowntoday
24

Write and review Playwright E2E tests for Langflow. An agent skill from langflow-ai/langflow.

langflow-ai/langflow156k—~3.3kAutomated safety check: PassMITtoday
25

Write, review, or upgrade Effect v4 code in the Composio CLI, cli-keyring, and json-schema-to-effect-schema packages, all pinned exactly to effect@4.0.0-rc.117 — Context.Service and explicit layers…

ComposioHQ/composio30k—~1kAutomated safety check: PassMITtoday
26

Disciplined diagnosis loop for hard bugs and performance regressions.

ywwynm/EverythingDone14418 repos~1.8kAutomated safety check: PassGPL-3.023 days ago
27

Sets the test-writing workflow for the repository: risk-first scenario lists, behavior-focused Vitest tests, a full run before each commit and a coverage target.

iOfficeAI/AionUi33k1 repo~1.2kAutomated safety check: PassApache-2.028 days ago
28

Runs live QA for the CodexBar app: provider usage matrix checks through its packaged CLI, config validation and menu checks, with 1Password-backed credentials handled safely.

steipete/CodexBar22k—~1.2kAutomated safety check: PassMITtoday
29

Connects an agent to a real Chrome instance through the Chrome DevTools MCP server, so it can inspect the DOM, read console errors and profile performance directly.

addyosmani/agent-skills102k4 repos~3.5kAutomated safety check: WarnMIT4 days ago
30

Triages and lands a batch of open Dependabot PRs in the Onyx repo, where main is gated exclusively by GitHub's merge queue: approves and enqueues green PRs, closes superseded duplicates, fixes…

onyx-dot-app/onyx32k1 repo~2.2kAutomated safety check: PassMITtoday
31

Analyze an Android pull request, branch, commit, or patch for user-visible changes and produce reproducible before/after screenshots from isolated builds.

permissionlesstech/bitchat-android7.7k—~2.6kAutomated safety check: PassGPL-3.0yesterday
32

Drives a real browser through the omowright library, either the user's own signed-in browser or a separate browser the code launches, for forms, QA, screenshots and scraping.

code-yeongyu/oh-my-openagent70k—~2.2kAutomated safety check: PassUnknowntoday
33

Score, evaluate, and iteratively improve any content or strategy using an auto-assembled panel of domain experts.

ericosiu/ai-marketing-skills3.6k2 repos~2.1kAutomated safety check: PassMIT15 days ago
34

Explains how to run agent integration tests against remote executors, using Docker for Linux or Wine for Windows, and how to opt tests in or skip them.

openinterpreter/openinterpreter69k2 repos~842Automated safety check: PassApache-2.0today
35

Starts a throwaway MongoDB 7.0 replica set and runs the Airbyte spec, check, discover and read commands against source-mongodb-v2 images for local end-to-end testing.

airbytehq/airbyte22k—~1.9kAutomated safety check: PassUnknownyesterday
36

Records a project's quality bar in CONSTRAINTS.md and watches diffs for signs an agent quietly weakened it, such as suppressions, skipped tests or lowered thresholds.

addyosmani/agent-skills102k2 repos~5.2kAutomated safety check: PassMIT4 days ago
37

Automates a Chrome or Chromium browser through the agent-browser CLI: navigate, fill forms, click, screenshot and extract data using element refs.

sipeed/picoclaw30k—~1.1kAutomated safety check: PassMIT13 days ago
38

Investigates a stubbornly failing Playwright test as a possible product bug, using error output, screenshots, traces and server code, and writes a structured bug report.

appsmithorg/appsmith41k—~1.5kAutomated safety check: PassApache-2.0yesterday
39

This skill should be used when the user is writing Go code and needs guidance on Go-specific pedantry: error wrapping with fmt.Errorf and %w, interface design (accept interfaces return structs)…

chromedp/chromedp13k—~3.7kAutomated safety check: PassMIT2 days ago
40

Verifies a delivery end to end by driving the real product on a CLI, web, desktop or iOS Simulator surface, capturing evidence and publishing a round with the lh CLI.

lobehub/lobehub83k—~9.7kAutomated safety check: PassApache-2.0today
41

Checks that every Python code block in a Markdown file actually compiles and runs, by extracting each block to a temporary file, executing it in an isolated subprocess, and writing a pass/fail…

google/adk-python22k—~1.4kAutomated safety check: PassApache-2.0today
42

Log genuine, recurring repository friction to .agents/PAPERCUTS.md — confusing setup, a flaky repo command or script, a misleading in-repo error, stale generated files, or a non-obvious gotcha that…

every-app/open-seo23k1 repo~1.2kAutomated safety check: PassMITyesterday
43

A skill your agent uses when the user mentions Yaak, a Yaak workspace, or the yaak command, or asks to call, hit, or smoke test HTTP/REST endpoints, save or organize API requests for reuse or manual…

mountain-loop/yaak19k—~1.9kAutomated safety check: PassMITtoday
44
44.TDD

Test-driven development. An agent skill from fossasia/eventyay-interpretation.

fossasia/eventyay-interpretation1.6k28 repos~1.1kAutomated safety check: PassApache-2.02 days ago
45

Run Kedro's local lint / format / type-check / tests on changed files (uses the project's pre-commit hooks, ruff, mypy, pytest, lint-imports, detect-secrets, Make targets — in the right venv), or…

kedro-org/kedro11k—~4kAutomated safety check: PassUnknownyesterday
46

Drives a web browser from the shell with the agent-browser CLI: open pages, read an element snapshot, click and fill by reference, grab text and screenshots.

nanocoai/nanoclaw31k3 repos~1.6kAutomated safety check: PassMITyesterday
47

Drives the Expo showcase app in an iOS simulator with agent-device to sweep every UI Kitten component in all theme and mapping combinations, reporting regressions with evidence.

akveo/react-native-ui-kitten11k—~2.3kAutomated safety check: PassMITyesterday
48

Creates a Java SDK end-to-end test for the Copilot SDK that runs against a recorded YAML snapshot through a replay proxy, so CI needs no real authentication.

github/copilot-sdk11k—~1.8kAutomated safety check: PassMITtoday

Questions, answered from the data.

What is the best testing and QA skill?

Agent Browser CLI (official) from vercel-labs/agent-browser ranks first of the 4,904 testing and QA skills listed here, with the highest score: its repository has 44k GitHub stars, 24 other GitHub owners carry a copy, its SKILL.md loads about 864 tokens and it passes the automated safety check with no findings. Next come Dogfood Exploratory QA and PR Babysitter.

Which testing and QA skills are official?

371 of the 4,904 testing and QA skills are official, published by the vendor's own GitHub organization: Agent Browser CLI, Dogfood Exploratory QA, Web Application Testing, Playwright CLI, Electron App Automation and 366 more.

How are these skills ranked?

By Skill Navigator score, which combines the GitHub stars of the skill's repository (shared across that repo's skills and discounted for large collections), how many other GitHub owners carry a copy of the skill, and automated SKILL.md quality checks, minus penalties for safety-check warnings and for each further skill from the same repository. Skills that fail the safety check are not listed.