Topic · Testing & QA
Best QA and bug reports skills for Claude Code, Codex and other agents.
- skills
- 559
- official
- 43
QA and bug reports skills, ranked
Ranked by score. Sort bymost stars,trending,newest,recently updated
| # | Skill | Repository | Stars | Used in | Tokens | Auto-check | Licence | Updated |
|---|---|---|---|---|---|---|---|---|
| 1 | Browser automation CLI for AI agents. Use when the user needs to interact with websites, including navigating pages, filling forms, clicking… | vercel-labs/ | 44k | 24 repos | ~864 | Automated safety check: Pass | Apache-2.0 | yesterday |
| 2 | Explores a web app with the agent-browser CLI to find bugs and UX problems, then writes a report with screenshots, repro videos and step-by-step reproduction for each issue. | vercel-labs/ | 44k | 8 repos | ~2.7k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 3 | Walks through an end-to-end smoke test of a DeerFlow deployment: pull the latest code, deploy with Docker or locally, verify services, run health checks and write a report. | bytedance/ | 83k | — | ~2.5k | Automated safety check: Notes | MIT | today |
| 4 | Runs live QA for the CodexBar app: provider usage matrix checks through its packaged CLI, config validation and menu checks, with 1Password-backed credentials handled safely. | steipete/ | 22k | — | ~1.2k | Automated safety check: Pass | MIT | today |
| 5 | Investigates a stubbornly failing Playwright test as a possible product bug, using error output, screenshots, traces and server code, and writes a structured bug report. | appsmithorg/ | 41k | — | ~1.5k | Automated safety check: Pass | Apache-2.0 | today |
| 6 | Verifies a delivery end to end by driving the real product on a CLI, web, desktop or iOS Simulator surface, capturing evidence and publishing a round with the lh CLI. | lobehub/ | 83k | — | ~9.7k | Automated safety check: Pass | Apache-2.0 | today |
| 7 | 7.Use Yaak A skill your agent uses when the user mentions Yaak, a Yaak workspace, or the yaak command, or asks to call, hit, or smoke test HTTP/REST endpoints, save or organize API requests for reuse or manual… | mountain-loop/ | 19k | — | ~1.9k | Automated safety check: Pass | MIT | today |
| 8 | Drives the Expo showcase app in an iOS simulator with agent-device to sweep every UI Kitten component in all theme and mapping combinations, reporting regressions with evidence. | akveo/ | 11k | — | ~2.3k | Automated safety check: Pass | MIT | yesterday |
| 9 | Tests the omo Codex plugin in an isolated CODEX_HOME with a local mock model, proving hooks fired through app-server notifications without touching ~/.codex. | code-yeongyu/ | 70k | — | ~1.9k | Automated safety check: Pass | Unknown | today |
| 10 | Creates GitHub issues for the current repository by choosing the matching issue template and following its format, with a permission check for engineering tasks. | CherryHQ/ | 52k | — | ~1.6k | Automated safety check: Pass | AGPL-3.0 | today |
| 11 | Drives a running OpenWork desktop window over CDP from the shell to evaluate JS, take screenshots, start sessions and send prompts for hand checks. | different-ai/ | 24k | — | ~465 | Automated safety check: Pass | Unknown | today |
| 12 | Turn a rough bug report, feature request, support note, or pull request into a short, plain-language issue focused on the problem and desired behavior. | every-app/ | 23k | 1 repo | ~1.2k | Automated safety check: Pass | MIT | today |
| 13 | Smoke test the Agent Builder feature branch end-to-end against a hermetic project scaffolded by the skill (linked to the current worktree). | mastra-ai/ | 29k | — | ~11k | Automated safety check: Notes | Unknown | today |
| 14 | Reproduce CUA-Harness experiments on WeaveBench from a GitHub checkout. | AMAP-ML/ | 1.7k | — | ~1.6k | Automated safety check: Pass | MIT | 1 mo ago |
| 15 | Proves a Codewhale change in the real product: a stamped release build, an atomic local install, fresh-shell verification and manual QA that automated gates cannot cover. | codewhale-hq/ | 41k | — | ~1.3k | Automated safety check: Pass | MIT | today |
| 16 | Makes the desktop app's model provider fail on demand, with refused connections, resets, stalls and HTTP 4xx and 5xx errors, so error and retry states can be reproduced. | different-ai/ | 24k | — | ~642 | Automated safety check: Pass | Unknown | today |
| 17 | Prepares the current DevSpace checkout or worktree for isolated local manual QA, covering QA state seeding, UI asset builds and snapshot resets. | Waishnav/ | 5.2k | — | ~440 | Automated safety check: Pass | MIT | 2 days ago |
| 18 | 18.Issue Fix Fix bugs in SkiaSharp C bindings. An agent skill from mono/SkiaSharp. | mono/ | 5.6k | — | ~5.1k | Automated safety check: Pass | MIT | today |
| 19 | Turns a confirmed MoviePilot bug or feature request into a structured upstream GitHub issue, but only after local diagnosis and an explicit request to file. | jxxghp/ | 12k | — | ~2.9k | Automated safety check: Pass | GPL-3.0 | today |
| 20 | Helps a reporter describe a bug, searches for duplicates and gathers diagnostic evidence kept separate from a short, human-worded issue draft. | OrchestratorInc/ | 13k | — | ~1.8k | Automated safety check: Pass | Apache-2.0 | today |
| 21 | 21.Pgjev Install, configure, query and explain pgjev (the jev PostgreSQL extension that filters, ranks and classifies rows with plain-language conditions via TypeSafe's Jev model). | realZachi/ | 1k | — | ~2.9k | Automated safety check: Pass | Unknown | yesterday |
| 22 | Records an annotated screen recording of the agent testing an app hands-on, then posts the video and a results summary to the PR and tracker issue. | michaelshimeles/ | 1.3k | 1 repo | ~3.9k | Automated safety check: Pass | No licence | 3 days ago |
| 23 | Smoke test Mastra projects locally or deploy to staging/production. | mastra-ai/ | 29k | — | ~4k | Automated safety check: Notes | Unknown | today |
| 24 | 24.Senior QA Comprehensive QA and testing skill for quality assurance, test automation, and testing strategies for ReactJS, NextJS, NodeJS applications. | nicepkg/ | 192 | 3 repos | ~1.1k | Automated safety check: Notes | No licence | 7 mo ago |
| 25 | Drives a real browser through Aside so the agent can open a page, read it, click through a flow, take screenshots and check console errors. | garrytan/ | 136k | — | ~8.1k | Automated safety check: Notes | MIT | today |
| 26 | Diagnose StaticPHP v3 failures. An agent skill from crazywhalecc/static-php-cli. | crazywhalecc/ | 1.9k | — | ~814 | Automated safety check: Pass | MIT | today |
| 27 | Tests the opencode coding agent itself: its CLI, server, plugin hooks and events, the terminal UI under tmux, and its SQLite session database, using tested helper scripts. | code-yeongyu/ | 70k | — | ~2.9k | Automated safety check: Pass | Unknown | today |
| 28 | Create GitHub issues using the gh CLI. An agent skill from NVIDIA/OpenShell. | NVIDIA/ | 15k | — | ~1.7k | Automated safety check: Pass | Apache-2.0 | today |
| 29 | Generate comprehensive test plans, manual test cases, regression test suites, and bug reports for QA engineers. | meshery/ | 151 | 5 repos | ~4.6k | Automated safety check: Pass | Apache-2.0 | 16 days ago |
| 30 | Smoke-tests a Skyvern deployment by checking the backend API, frontend rendering, browser session provisioning and workflow execution in sequence. | Skyvern-AI/ | 23k | — | ~1.5k | Automated safety check: Pass | AGPL-3.0 | today |
| 31 | Fires known chat states in the running OpenWork desktop app, such as provider errors, retries and tool steps, so you can check how each renders. | different-ai/ | 24k | — | ~673 | Automated safety check: Pass | Unknown | today |
| 32 | Complete agent-ready workflow for Qwen-family MTP or nextn GGUF conversion and release. | R6410418/ | 1.7k | — | ~1.7k | Automated safety check: Pass | MIT | 2 mo ago |
| 33 | Use the host-side agent-browser CLI for local browser smoke tests, screenshots, snapshots, and simple UI validation against forwarded localhost URLs. | superagent-ai/ | 3.5k | — | ~633 | Automated safety check: Pass | MIT | 3 mo ago |
| 34 | Tests a SwiftUI app on a real iPhone connected by USB, reading the Swift source and then looping through screenshot, analysis and action to find bugs. | garrytan/ | 136k | — | ~10k | Automated safety check: Notes | MIT | today |
| 35 | Plays an authorized Game Boy or Game Boy Color ROM in one persistent headless Coffee GB session, inspecting frames and keeping an action trace for replay or tests. | trekawek/ | 1.2k | — | ~1.3k | Automated safety check: Pass | MIT | 6 days ago |
| 36 | Uses a clean Parallels macOS VM to test GUI automation, TCC permission prompts and screenshot tools like Peekaboo, verifying results from outside the guest. | steipete/ | 7.3k | — | ~1.8k | Automated safety check: Pass | MIT | 2 days ago |
| 37 | A skill your agent uses when starting or verifying Hermes Agent CN Desktop with the latest Hermes-CN-Desktop and Hermes-CN-Core branches — dev smoke test (pnpm tauri:dev), packaged beta/release… | Eynzof/ | 1.7k | — | ~1.8k | Automated safety check: Pass | Unknown | 16 days ago |
| 38 | Analyzes a single GitHub issue at a time. An agent skill from ClickHouse/clickhouse-java. | ClickHouse/ | 1.6k | — | ~904 | Automated safety check: Pass | Apache-2.0 | yesterday |
| 39 | A skill your agent uses when working in a repo with agentacct MCP configured, or when asked to track coding-agent work, smoke-test agentacct integrations, or report objective AI-agent task evidence. | mikehasa/ | 765 | — | ~1.6k | Automated safety check: Pass | MIT | 4 days ago |
| 40 | Create structured Jira tickets for Dynamo from bug reports, failing tests, or feature requests. | DynamoDS/ | 2k | — | ~1.1k | Automated safety check: Pass | Apache-2.0 | today |
| 41 | Rigor Run skill for README-first deep learning repo reproduction. | lllllllama/ | 497 | 2 repos | ~691 | Automated safety check: Pass | MIT | 14 days ago |
| 42 | Test IronRDP ActiveX in the real mstsc.exe child launched by MsRdpEx, including the native credential bridge, shell behavior, resize handling, clipboard startup, and bounded UI stress. | Devolutions/ | 3.2k | — | ~1.3k | Automated safety check: Pass | Apache-2.0 | today |
| 43 | Checks changes to the senpi coding agent by driving the real CLI from source in an isolated sandbox, over RPC, terminal UI, mock model and CLI smoke channels. | code-yeongyu/ | 470 | — | ~2.7k | Automated safety check: Notes | MIT | today |
| 44 | Review and validate claims using counter-hypothesis testing. | bradygaster/ | 3.3k | — | ~503 | Automated safety check: Pass | MIT | yesterday |
| 45 | Create and triage GitHub issues from repository evidence. An agent skill from Gentleman-Programming/gentle-shell. | Gentleman-Programming/ | 1.2k | — | ~2.5k | Automated safety check: Pass | Apache-2.0 | today |
| 46 | Confirms that a change to PlotJuggler 4 really works in the running app by proving the rebuild, launching with real data and measuring the result. | PlotJuggler/ | 6.2k | — | ~956 | Automated safety check: Pass | MPL-2.0 | 6 days ago |
| 47 | 47.Ya Run Launch and drive Yep Anywhere — an isolated dev server plus real browser interaction — and run the repository's check suite. | kzahel/ | 534 | — | ~957 | Automated safety check: Pass | MIT | today |
| 48 | Build, validate, and run the claude-osint skills repo — check SKILL.md frontmatter, run the secretscan.py and h1reference.py helpers, run sync-skill-content.sh, run the smoke test. | elementalsouls/ | 2.8k | — | ~1.2k | Automated safety check: Pass | MIT | 1 mo ago |
Questions, answered from the data.
What is the best QA and bug reports skill?
Agent Browser CLI (official) from vercel-labs/agent-browser ranks first of the 559 QA and bug reports skills listed here, with the highest score: its repository has 44k GitHub stars, 24 other GitHub owners carry a copy, its SKILL.md loads about 864 tokens and it passes the automated safety check with no findings. Next come Dogfood Exploratory QA and DeerFlow Smoke Test.
Which QA and bug reports skills are official?
43 of the 559 QA and bug reports skills are official, published by the vendor's own GitHub organization: Agent Browser CLI, Dogfood Exploratory QA, Create GitHub Issue, Issues Deduplication, Fix Analyzer Bug and 38 more.
How are these skills ranked?
By Skill Navigator score, which combines the GitHub stars of the skill's repository (shared across that repo's skills and discounted for large collections), how many other GitHub owners carry a copy of the skill, and automated SKILL.md quality checks, minus penalties for safety-check warnings and for each further skill from the same repository. Skills that fail the safety check are not listed.