Topic · Testing & QA
Best quality gates skills for Claude Code, Codex and other agents.
- skills
- 342
- official
- 6
Quality gates skills, ranked
Ranked by score. Sort bymost stars,trending,newest,recently updated
| # | Skill | Repository | Stars | Used in | Tokens | Auto-check | Licence | Updated |
|---|---|---|---|---|---|---|---|---|
| 1 | Reads the state and results of Caveman Cloud experiments and reports one recommendation or a block, without changing an experiment's lifecycle itself. | JuliusBrussee/ | 110k | 1 repo | ~975 | Automated safety check: Pass | Apache-2.0 | today |
| 2 | Evaluates an Agent Skill bundle before release for structure, trigger quality, artifact improvement, script correctness, safety, installed-tree integrity and host portability. | rohitg00/ | 65k | — | ~1k | Automated safety check: Pass | MIT | today |
| 3 | Builds new skills from scattered docs, specs or code, and refactors existing skills for clear triggers, reliable activation and a quality checklist. | 2025Emma/ | 23k | 2 repos | ~2k | Automated safety check: Pass | MIT | 9 mo ago |
| 4 | Sets the test-writing workflow for the repository: risk-first scenario lists, behavior-focused Vitest tests, a full run before each commit and a coverage target. | iOfficeAI/ | 33k | 1 repo | ~1.2k | Automated safety check: Pass | Apache-2.0 | 28 days ago |
| 5 | Score, evaluate, and iteratively improve any content or strategy using an auto-assembled panel of domain experts. | ericosiu/ | 3.6k | 2 repos | ~2.1k | Automated safety check: Pass | MIT | 14 days ago |
| 6 | Meta-skill that turns docs, APIs, code or specs into a reusable skill with references and a quality gate, and refactors skills that are unclear or misfire. | tradecatlabs/ | 17k | 1 repo | ~2.4k | Automated safety check: Pass | MIT | 6 days ago |
| 7 | Records a project's quality bar in CONSTRAINTS.md and watches diffs for signs an agent quietly weakened it, such as suppressions, skipped tests or lowered thresholds. | addyosmani/ | 102k | 2 repos | ~5.2k | Automated safety check: Pass | MIT | 4 days ago |
| 8 | Drives Android devices through SoloPi's typed command line to record, replay and verify app behavior, with device pools and signed on-device decision models. | alipay/ | 6.3k | — | ~3.6k | Automated safety check: Pass | Apache-2.0 | 1 mo ago |
| 9 | Decides from local git changes whether make bazel_prepare must run in TiDB before build or test commands, and reports the evidence for the decision. | pingcap/ | 41k | — | ~408 | Automated safety check: Pass | Apache-2.0 | today |
| 10 | 10.Quality Scan Runs a scoped, read-only quality scan on a repository candidate and reports exact evidence, failures and frozen debt, without treating a static scan as release acceptance. | diegosouzapw/ | 74k | — | ~748 | Automated safety check: Pass | MIT | today |
| 11 | 11.Test Guard Reviews newly written or edited tests against nine rules that cut test bloat, such as mock-heavy checks and near-duplicate cases, before they are committed. | amElnagdy/ | 1.3k | 2 repos | ~2.1k | Automated safety check: Pass | MIT | 3 mo ago |
| 12 | Inspect, review, or apply authorized triage decisions to SonarCloud issues and security hotspots; also review the Sonar helpers. | netdata/ | 81k | — | ~2.8k | Automated safety check: Notes | GPL-3.0 | today |
| 13 | 13.Evaluation This skill should be used when building agent evaluation systems: deterministic checks, regression suites, multi-dimensional rubrics, quality gates, production monitoring, baseline comparison, and… | guanyang/ | 973 | 2 repos | ~4.2k | Automated safety check: Pass | MIT | today |
| 14 | Runs Cherry Studio's critical-path regression suite as deterministic Playwright E2E tests through a GitHub workflow on macOS and Windows runners. | CherryHQ/ | 52k | — | ~1.2k | Automated safety check: Pass | AGPL-3.0 | today |
| 15 | Audits game assets against naming conventions, file size budgets and format standards, and finds orphaned assets and missing references. | Donchitos/ | 26k | — | ~2k | Automated safety check: Pass | MIT | 8 days ago |
| 16 | Convert PRDs to prd.json format for ralph-tui execution. An agent skill from subsy/ralph-tui. | subsy/ | 2.5k | 1 repo | ~2.6k | Automated safety check: Pass | MIT | 1 mo ago |
| 17 | Scans a module directory for the required README.md and DESIGN.md plus recommended files and reports what is missing, so a module is not delivered incomplete. | fengshao1227/ | 5.9k | — | ~473 | Automated safety check: Notes | MIT | 22 days ago |
| 18 | Plans and audits Day-0 SGLang support for a new model release: scope, architecture gaps, PR order, validation gates and sanitized public evidence. | BBuf/ | 900 | — | ~2.3k | Automated safety check: Pass | No licence | 2 days ago |
| 19 | Runs a gated finish-line checklist before committing a PlotJuggler PJ4 change: build proof, red-test triage, hooks, docs freshness and a diff self-review. | PlotJuggler/ | 6.2k | — | ~1.3k | Automated safety check: Pass | MPL-2.0 | 6 days ago |
| 20 | Reviews BiSheng spec, design and tasks documents with checklists for PRD gaps, handover readiness and acceptance traceability, producing a report or an LGTM. | dataelement/ | 12k | — | ~717 | Automated safety check: Pass | Apache-2.0 | 7 days ago |
| 21 | Critically review strategy drafts from edge-strategy-designer for edge plausibility, overfitting risk, sample size adequacy, and execution realism. | tradermonty/ | 3k | 1 repo | ~988 | Automated safety check: Pass | MIT | yesterday |
| 22 | Designs and verifies a deterministic grader that measures whether a GitHub Agentic Workflow run reached its real-world or repository outcome. | github/ | 5.3k | — | ~6.8k | Automated safety check: Pass | MIT | today |
| 23 | Wrap up the current session: verify quality gate passed, remind user to commit, archive completed tasks, and record session progress to the developer journal. | ROYIANS/ | 135 | 5 repos | ~917 | Automated safety check: Pass | MIT | 1 mo ago |
| 24 | Pre-commit review: security scan, quality gates, auto-fix. An agent skill from HezaoHezao/poirot. | HezaoHezao/ | 250 | 5 repos | ~1.6k | Automated safety check: Pass | MIT | 2 mo ago |
| 25 | Automates CI/CD pipeline setup. An agent skill from dzhalaevd/Donatello. | dzhalaevd/ | 135 | 7 repos | ~2.7k | Automated safety check: Notes | Apache-2.0 | 4 days ago |
| 26 | 26.Sonarqube Operate SonarQube-enabled repositories through the SonarQube CLI (sonar): verify authentication, discover project keys, inspect project metadata, issues, measures, and quality gates, analyze changed… | DougTrajano/ | 377 | — | ~2.2k | Automated safety check: Pass | MIT | 3 days ago |
| 27 | Runs a light convention check on one finished spec-driven task, choosing checks by task type and ending in pass, pass-with-notes or needs-fix. | dataelement/ | 12k | — | ~652 | Automated safety check: Pass | Apache-2.0 | 7 days ago |
| 28 | Generate or convert Claude Code prompt files — command orchestrators, skill files, agent role definitions, or style conversion of existing files. | catlog22/ | 2.1k | 1 repo | ~4.7k | Automated safety check: Notes | MIT | 3 mo ago |
| 29 | Inspect SonarQube Cloud/SonarCloud findings for this repository using local .env credentials. | hbmartin/ | 275 | — | ~508 | Automated safety check: Notes | GPL-3.0 | 2 mo ago |
| 30 | Primary/default skill for UI design, product design, web design, landing pages, dashboards, product screens, redesigns, visual polish, frontend/CSS styling, design systems, components, responsive… | referodesign/ | 292 | — | ~5.3k | Automated safety check: Pass | MIT | 1 mo ago |
| 31 | Adds a release eval scorecard to a GAIA hub agent by writing a harness adapter, running a real eval, and wiring the result into the agent's README and release gate. | amd/ | 1.6k | — | ~2.6k | Automated safety check: Pass | MIT | today |
| 32 | 32.Pull Request Take a change from working tree to a merge-ready pull request, then keep iterating until the CI checks and AI reviewers (CodeRabbit, cubic) all pass. | noh-rs/ | 156 | — | ~4.1k | Automated safety check: Pass | MIT | 5 days ago |
| 33 | Runs a staged pull request or diff review using selected reviewer personas, checking the change against its stated intent and project standards before producing findings. | EveryInc/ | 25k | — | ~2k | Automated safety check: Pass | MIT | today |
| 34 | Turn a research brief, finished script, or generated MP4 into a reviewed and verified vertical explainer video. | runesleo/ | 119 | — | ~2.2k | Automated safety check: Pass | MIT | 10 days ago |
| 35 | Turn arbitrary source content (README, article, story, slides, deck, data/report, product description, tutorial text, audio/transcript, or a bare topic) into a finished, high-quality MP4 video. | architectds/ | 117 | — | ~2.4k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 36 | Disciplined task execution with planning, verification, and self-improvement loops. | vxcozy/ | 115 | — | ~1k | Automated safety check: Pass | MIT | 5 mo ago |
| 37 | Creates phase-based feature plans with quality gates and incremental delivery structure. | serendipity1004/ | 176 | — | ~2.4k | Automated safety check: Pass | No licence | 9 mo ago |
| 38 | Reference for MoAI-ADK's core development principles: TRUST 5 quality gates, SPEC-first domain-driven workflow, agent delegation and token budgeting. | modu-ai/ | 1.2k | — | ~5k | Automated safety check: Pass | Apache-2.0 | today |
| 39 | Audits version freshness, French and English parity and metadata of whitepapers and recap cards in the claude-code-ultimate-guide repo, with optional fix suggestions. | FlorianBruniaux/ | 6.1k | — | ~3.8k | Automated safety check: Pass | CC-BY-SA-4.0 | yesterday |
| 40 | 40.Nelson Orchestrates multi-agent task execution using a Royal Navy squadron metaphor — from mission planning through parallel work coordination to stand-down. | Aspegio/ | 421 | — | ~11k | Automated safety check: Warn | MIT | 3 mo ago |
| 41 | Image-to-code replication pipeline. An agent skill from Yu-369/VibeCurb. | Yu-369/ | 980 | — | ~8.7k | Automated safety check: Pass | MIT | 2 mo ago |
| 42 | Runs local checks and a repo-reviewer subagent over the whole branch diff before a pull request is opened, then records the approval in the PR description. | yuga-hashimoto/ | 123 | — | ~710 | Automated safety check: Pass | MIT | today |
| 43 | Wrap up the current session: verify quality gate passed, archive completed tasks, record session progress, push branch, and create PR. | OrtonY/ | 122 | — | ~1.3k | Automated safety check: Pass | Apache-2.0 | 3 mo ago |
| 44 | Applies Procoder's senior-developer discipline in a repository: run the commit gate, format through the binary and work through specs, plans and todos. | azrtydxb/ | 211 | — | ~3.7k | Automated safety check: Pass | Apache-2.0 | 9 days ago |
| 45 | A required checklist for after code changes: validate each changed layer, check that docs and other layers stay in sync, and test before committing or opening a PR. | ZeroDeng01/ | 1.7k | — | ~4.4k | Automated safety check: Pass | MIT | 2 days ago |
| 46 | 46.Bump Bump Houndarr version and prepare a release PR. An agent skill from av1155/houndarr. | av1155/ | 292 | — | ~1.2k | Automated safety check: Pass | AGPL-3.0 | 2 days ago |
| 47 | Adds and debugs JacRed torrent/anime tracker parsers (site probe, auth, magnet policy, SyncService, cron, fixtures, dry-run, OpenAPI wiring). | jacred-fdb/ | 124 | — | ~1.4k | Automated safety check: Pass | AGPL-3.0 | 2 days ago |
| 48 | Runs local checks on maintainer docs for stale review dates, missing manifest paths and plan state drift, matching what the CI workflow enforces. | liaohch3/ | 3.3k | — | ~561 | Automated safety check: Pass | MIT | 15 days ago |
Questions, answered from the data.
What is the best quality gates skill?
Caveman Experiment Manager from JuliusBrussee/caveman ranks first of the 342 quality gates skills listed here, with the highest score: its repository has 110k GitHub stars, 1 other GitHub owner carry a copy, its SKILL.md loads about 975 tokens and it passes the automated safety check with no findings. Next come Skill Release Gate and Skill Authoring and Refactoring.
Which quality gates skills are official?
6 of the 342 quality gates skills are official, published by the vendor's own GitHub organization: Operational Value Designer, Verify Tests Catch the Bug, gh-aw Developer Rules, Sonar Quality Gate, Audit Integrity and 1 more.
How are these skills ranked?
By Skill Navigator score, which combines the GitHub stars of the skill's repository (shared across that repo's skills and discounted for large collections), how many other GitHub owners carry a copy of the skill, and automated SKILL.md quality checks, minus penalties for safety-check warnings and for each further skill from the same repository. Skills that fail the safety check are not listed.