Topic · Testing & QA

Best QA and bug reports skills for Claude Code, Codex and other agents.

Skills that structure manual QA, exploratory testing and clear bug reports.
skills
559
official
43

QA and bug reports skills, ranked

Ranked by score. Sort bymost stars,trending,newest,recently updated

QA and bug reports skills, ranked
#SkillRepositoryStarsUsed inTokensAuto-checkLicenceUpdated
1

Browser automation CLI for AI agents. Use when the user needs to interact with websites, including navigating pages, filling forms, clicking…

vercel-labs/agent-browser44k24 repos~864Automated safety check: PassApache-2.0yesterday
2

Explores a web app with the agent-browser CLI to find bugs and UX problems, then writes a report with screenshots, repro videos and step-by-step reproduction for each issue.

vercel-labs/agent-browser44k8 repos~2.7kAutomated safety check: PassApache-2.0yesterday
3

Walks through an end-to-end smoke test of a DeerFlow deployment: pull the latest code, deploy with Docker or locally, verify services, run health checks and write a report.

bytedance/deer-flow83k—~2.5kAutomated safety check: NotesMITtoday
4

Runs live QA for the CodexBar app: provider usage matrix checks through its packaged CLI, config validation and menu checks, with 1Password-backed credentials handled safely.

steipete/CodexBar22k—~1.2kAutomated safety check: PassMITtoday
5

Investigates a stubbornly failing Playwright test as a possible product bug, using error output, screenshots, traces and server code, and writes a structured bug report.

appsmithorg/appsmith41k—~1.5kAutomated safety check: PassApache-2.0today
6

Verifies a delivery end to end by driving the real product on a CLI, web, desktop or iOS Simulator surface, capturing evidence and publishing a round with the lh CLI.

lobehub/lobehub83k—~9.7kAutomated safety check: PassApache-2.0today
7

A skill your agent uses when the user mentions Yaak, a Yaak workspace, or the yaak command, or asks to call, hit, or smoke test HTTP/REST endpoints, save or organize API requests for reuse or manual…

mountain-loop/yaak19k—~1.9kAutomated safety check: PassMITtoday
8

Drives the Expo showcase app in an iOS simulator with agent-device to sweep every UI Kitten component in all theme and mapping combinations, reporting regressions with evidence.

akveo/react-native-ui-kitten11k—~2.3kAutomated safety check: PassMITyesterday
9

Tests the omo Codex plugin in an isolated CODEX_HOME with a local mock model, proving hooks fired through app-server notifications without touching ~/.codex.

code-yeongyu/oh-my-openagent70k—~1.9kAutomated safety check: PassUnknowntoday
10

Creates GitHub issues for the current repository by choosing the matching issue template and following its format, with a permission check for engineering tasks.

CherryHQ/cherry-studio52k—~1.6kAutomated safety check: PassAGPL-3.0today
11

Drives a running OpenWork desktop window over CDP from the shell to evaluate JS, take screenshots, start sessions and send prompts for hand checks.

different-ai/openwork24k—~465Automated safety check: PassUnknowntoday
12

Turn a rough bug report, feature request, support note, or pull request into a short, plain-language issue focused on the problem and desired behavior.

every-app/open-seo23k1 repo~1.2kAutomated safety check: PassMITtoday
13

Smoke test the Agent Builder feature branch end-to-end against a hermetic project scaffolded by the skill (linked to the current worktree).

mastra-ai/mastra29k—~11kAutomated safety check: NotesUnknowntoday
14

Reproduce CUA-Harness experiments on WeaveBench from a GitHub checkout.

AMAP-ML/LongHorizon-Harness1.7k—~1.6kAutomated safety check: PassMIT1 mo ago
15

Proves a Codewhale change in the real product: a stamped release build, an atomic local install, fresh-shell verification and manual QA that automated gates cannot cover.

codewhale-hq/Codewhale41k—~1.3kAutomated safety check: PassMITtoday
16

Makes the desktop app's model provider fail on demand, with refused connections, resets, stalls and HTTP 4xx and 5xx errors, so error and retry states can be reproduced.

different-ai/openwork24k—~642Automated safety check: PassUnknowntoday
17

Prepares the current DevSpace checkout or worktree for isolated local manual QA, covering QA state seeding, UI asset builds and snapshot resets.

Waishnav/devspace5.2k—~440Automated safety check: PassMIT2 days ago
18

Fix bugs in SkiaSharp C bindings. An agent skill from mono/SkiaSharp.

mono/SkiaSharp5.6k—~5.1kAutomated safety check: PassMITtoday
19

Turns a confirmed MoviePilot bug or feature request into a structured upstream GitHub issue, but only after local diagnosis and an explicit request to file.

jxxghp/MoviePilot12k—~2.9kAutomated safety check: PassGPL-3.0today
20

Helps a reporter describe a bug, searches for duplicates and gathers diagnostic evidence kept separate from a short, human-worded issue draft.

OrchestratorInc/agent-orchestrator13k—~1.8kAutomated safety check: PassApache-2.0today
21

Install, configure, query and explain pgjev (the jev PostgreSQL extension that filters, ranks and classifies rows with plain-language conditions via TypeSafe's Jev model).

realZachi/pg-jev1k—~2.9kAutomated safety check: PassUnknownyesterday
22

Records an annotated screen recording of the agent testing an app hands-on, then posts the video and a results summary to the PR and tracker issue.

michaelshimeles/skills1.3k1 repo~3.9kAutomated safety check: PassNo licence3 days ago
23

Smoke test Mastra projects locally or deploy to staging/production.

mastra-ai/mastra29k—~4kAutomated safety check: NotesUnknowntoday
24

Comprehensive QA and testing skill for quality assurance, test automation, and testing strategies for ReactJS, NextJS, NodeJS applications.

nicepkg/auto-company1923 repos~1.1kAutomated safety check: NotesNo licence7 mo ago
25

Drives a real browser through Aside so the agent can open a page, read it, click through a flow, take screenshots and check console errors.

garrytan/gstack136k—~8.1kAutomated safety check: NotesMITtoday
26

Diagnose StaticPHP v3 failures. An agent skill from crazywhalecc/static-php-cli.

crazywhalecc/static-php-cli1.9k—~814Automated safety check: PassMITtoday
27

Tests the opencode coding agent itself: its CLI, server, plugin hooks and events, the terminal UI under tmux, and its SQLite session database, using tested helper scripts.

code-yeongyu/oh-my-openagent70k—~2.9kAutomated safety check: PassUnknowntoday
28

Create GitHub issues using the gh CLI. An agent skill from NVIDIA/OpenShell.

NVIDIA/OpenShell15k—~1.7kAutomated safety check: PassApache-2.0today
29

Generate comprehensive test plans, manual test cases, regression test suites, and bug reports for QA engineers.

meshery/meshery-operator1515 repos~4.6kAutomated safety check: PassApache-2.016 days ago
30

Smoke-tests a Skyvern deployment by checking the backend API, frontend rendering, browser session provisioning and workflow execution in sequence.

Skyvern-AI/skyvern23k—~1.5kAutomated safety check: PassAGPL-3.0today
31

Fires known chat states in the running OpenWork desktop app, such as provider errors, retries and tool steps, so you can check how each renders.

different-ai/openwork24k—~673Automated safety check: PassUnknowntoday
32

Complete agent-ready workflow for Qwen-family MTP or nextn GGUF conversion and release.

R6410418/Jackrong-llm-finetuning-guide1.7k—~1.7kAutomated safety check: PassMIT2 mo ago
33

Use the host-side agent-browser CLI for local browser smoke tests, screenshots, snapshots, and simple UI validation against forwarded localhost URLs.

superagent-ai/grok-cli3.5k—~633Automated safety check: PassMIT3 mo ago
34

Tests a SwiftUI app on a real iPhone connected by USB, reading the Swift source and then looping through screenshot, analysis and action to find bugs.

garrytan/gstack136k—~10kAutomated safety check: NotesMITtoday
35

Plays an authorized Game Boy or Game Boy Color ROM in one persistent headless Coffee GB session, inspecting frames and keeping an action trace for replay or tests.

trekawek/coffee-gb1.2k—~1.3kAutomated safety check: PassMIT6 days ago
36

Uses a clean Parallels macOS VM to test GUI automation, TCC permission prompts and screenshot tools like Peekaboo, verifying results from outside the guest.

steipete/agent-scripts7.3k—~1.8kAutomated safety check: PassMIT2 days ago
37

A skill your agent uses when starting or verifying Hermes Agent CN Desktop with the latest Hermes-CN-Desktop and Hermes-CN-Core branches — dev smoke test (pnpm tauri:dev), packaged beta/release…

Eynzof/Hermes-CN-Desktop1.7k—~1.8kAutomated safety check: PassUnknown16 days ago
38

Analyzes a single GitHub issue at a time. An agent skill from ClickHouse/clickhouse-java.

ClickHouse/clickhouse-java1.6k—~904Automated safety check: PassApache-2.0yesterday
39

A skill your agent uses when working in a repo with agentacct MCP configured, or when asked to track coding-agent work, smoke-test agentacct integrations, or report objective AI-agent task evidence.

mikehasa/agentacct765—~1.6kAutomated safety check: PassMIT4 days ago
40

Create structured Jira tickets for Dynamo from bug reports, failing tests, or feature requests.

DynamoDS/Dynamo2k—~1.1kAutomated safety check: PassApache-2.0today
41

Rigor Run skill for README-first deep learning repo reproduction.

lllllllama/RigorPilot-Skills4972 repos~691Automated safety check: PassMIT14 days ago
42

Test IronRDP ActiveX in the real mstsc.exe child launched by MsRdpEx, including the native credential bridge, shell behavior, resize handling, clipboard startup, and bounded UI stress.

Devolutions/IronRDP3.2k—~1.3kAutomated safety check: PassApache-2.0today
43

Checks changes to the senpi coding agent by driving the real CLI from source in an isolated sandbox, over RPC, terminal UI, mock model and CLI smoke channels.

code-yeongyu/senpi470—~2.7kAutomated safety check: NotesMITtoday
44

Review and validate claims using counter-hypothesis testing.

bradygaster/squad3.3k—~503Automated safety check: PassMITyesterday
45

Create and triage GitHub issues from repository evidence. An agent skill from Gentleman-Programming/gentle-shell.

Gentleman-Programming/gentle-shell1.2k—~2.5kAutomated safety check: PassApache-2.0today
46

Confirms that a change to PlotJuggler 4 really works in the running app by proving the rebuild, launching with real data and measuring the result.

PlotJuggler/PlotJuggler6.2k—~956Automated safety check: PassMPL-2.06 days ago
47

Launch and drive Yep Anywhere — an isolated dev server plus real browser interaction — and run the repository's check suite.

kzahel/yepanywhere534—~957Automated safety check: PassMITtoday
48

Build, validate, and run the claude-osint skills repo — check SKILL.md frontmatter, run the secretscan.py and h1reference.py helpers, run sync-skill-content.sh, run the smoke test.

elementalsouls/Claude-OSINT2.8k—~1.2kAutomated safety check: PassMIT1 mo ago

Questions, answered from the data.

What is the best QA and bug reports skill?

Agent Browser CLI (official) from vercel-labs/agent-browser ranks first of the 559 QA and bug reports skills listed here, with the highest score: its repository has 44k GitHub stars, 24 other GitHub owners carry a copy, its SKILL.md loads about 864 tokens and it passes the automated safety check with no findings. Next come Dogfood Exploratory QA and DeerFlow Smoke Test.

Which QA and bug reports skills are official?

43 of the 559 QA and bug reports skills are official, published by the vendor's own GitHub organization: Agent Browser CLI, Dogfood Exploratory QA, Create GitHub Issue, Issues Deduplication, Fix Analyzer Bug and 38 more.

How are these skills ranked?

By Skill Navigator score, which combines the GitHub stars of the skill's repository (shared across that repo's skills and discounted for large collections), how many other GitHub owners carry a copy of the skill, and automated SKILL.md quality checks, minus penalties for safety-check warnings and for each further skill from the same repository. Skills that fail the safety check are not listed.