Developer tool
pytest agent skills for Claude Code, Codex and other agents.
- skills
- 268
- official
- 29
- Type
- Developer tool
- Website
- docs.pytest.org
- Official GitHub
- pytest-dev
pytest skills, ranked
Ranked by score. Sort bymost stars,trending,newest,recently updated
Official (29 skills)
| # | Skill | Repository | Stars | Used in | Tokens | Auto-check | Licence | Updated |
|---|---|---|---|---|---|---|---|---|
| 1 | Checks that every Python code block in a Markdown file actually compiles and runs, by extracting each block to a temporary file, executing it in an isolated subprocess, and writing a pass/fail… | google/ | 22k | — | ~1.4k | Automated safety check: Pass | Apache-2.0 | today |
| 2 | Sets up a local ADK Python development environment in a git clone of the open-source adk-python repository: a uv virtual environment, all dependency extras, pre-commit hooks, and a first unit-test… | google/ | 22k | — | ~993 | Automated safety check: Notes | Apache-2.0 | today |
| 3 | Runs and debugs ONNX Runtime tests: Google Test executables for C++ and unittest or pytest for Python, with filters and build-directory guidance. | microsoft/ | 22k | — | ~1.8k | Automated safety check: Pass | MIT | today |
| 4 | Builds ADK (Agent Development Kit) Python agents: LLM agents with tools, graph workflows of function and agent nodes, conditional routing, fan-out and join, schema-validated delegation between… | google/ | 22k | — | ~879 | Automated safety check: Pass | Apache-2.0 | today |
| 5 | Runs the ONNX Runtime transformers Python tests against a GPU wheel and proves the cuDNN flash attention path was used rather than a silent fallback. | microsoft/ | 22k | — | ~2.9k | Automated safety check: Pass | MIT | today |
| 6 | Draft and revise concise, human-focused GitHub issues for pytest-kind-ng. | NVIDIA/ | 309 | — | ~1.1k | Automated safety check: Pass | Apache-2.0 | 1 mo ago |
| 7 | Guides writing or refactoring tests for the Mistral Vibe CLI agent so they check behavior through stable boundaries, like tool invocation or saved session shape, instead of internal calls. | mistralai/ | 5.1k | — | ~1.2k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 8 | Guide to the Megatron-LM test system: layout, recipe YAML, running and adding unit and functional tests, golden values, marker filters and CI parity. | NVIDIA/ | 18k | — | ~2.1k | Automated safety check: Pass | Apache-2.0 | today |
| 9 | Recover from GPU-busy / GPU-unavailable failures. An agent skill from facebookexperimental/triton. | facebookexperimental/ | 201 | — | ~709 | Automated safety check: Pass | MIT | today |
| 10 | Coding and testing rules for Mistral Vibe so code works on Linux, macOS and Windows even though CI runs the test suite on Linux only. | mistralai/ | 5.1k | — | ~551 | Automated safety check: Pass | Apache-2.0 | yesterday |
| 11 | Analyze assertion quality, depth, variety, and false confidence in existing tests. | dotnet/ | 5.6k | 1 repo | ~4.8k | Automated safety check: Pass | MIT | today |
| 12 | ALWAYS USE whenever asked to write, add, or generate unit tests for existing code in xUnit, MSTest, NUnit, pytest, Vitest/Jest, Go, or another framework, including "tests only for" one helper… | dotnet/ | 5.6k | 1 repo | ~4.7k | Automated safety check: Pass | MIT | today |
| 13 | Sets up Python projects and standalone scripts with uv, ruff, ty, pytest and prek, and helps move existing projects off pip, Poetry, mypy and black. | trailofbits/ | 7.4k | — | ~2.5k | Automated safety check: Pass | CC-BY-SA-4.0 | 5 days ago |
| 14 | Write tests for bocpy Behavior-Oriented Concurrency code. An agent skill from microsoft/bocpy. | microsoft/ | 200 | — | ~6.1k | Automated safety check: Pass | MIT | 9 days ago |
| 15 | Provides file paths to language-specific reference files for the test ANALYSIS skills (assertion-quality, test-anti-patterns, test-gap-analysis, test-smell-detection, test-tagging). | microsoft/ | 1k | 2 repos | ~1.1k | Automated safety check: Pass | MIT | today |
| 16 | Run pytest tests with coverage, discover lines missing coverage, and increase coverage to 100%. | github/ | 40k | 4 repos | ~282 | Automated safety check: Pass | MIT | today |
| 17 | Analyzes the variety and depth of assertions across test suites in any language. | microsoft/ | 1k | — | ~4.1k | Automated safety check: Pass | MIT | today |
| 18 | Generates and writes new unit tests for any programming language — scaffolds .NET test projects, pytest suites, Vitest/Jest suites, Go test files, and JUnit suites, and configures coverage tooling… | microsoft/ | 1k | — | ~2.7k | Automated safety check: Pass | MIT | today |
| 19 | Performs pseudo-mutation analysis on production code in any language to find gaps in existing test suites. | microsoft/ | 1k | — | ~4k | Automated safety check: Pass | MIT | today |
| 20 | Deep-dive audit using the full testsmells.org 19-smell academic catalog for tests in any language. | microsoft/ | 1k | — | ~5.3k | Automated safety check: Pass | MIT | today |
| 21 | Analyzes test suites in any language and tags each test with a standardized set of traits (positive, negative, critical-path, boundary, smoke, regression, integration, performance, security). | microsoft/ | 1k | — | ~4.3k | Automated safety check: Pass | MIT | today |
| 22 | Audits existing test code in any language for anti-patterns and quality issues — produces a severity-ranked report (Critical / Warning / Info) with concrete code-level fixes. | microsoft/ | 1k | — | ~4.9k | Automated safety check: Pass | MIT | today |
| 23 | Create Earth2Studio diagnostic model wrappers for single-step data transformations, including simple derived diagnostics, packaged AutoModel diagnostics, and generative or diffusion diagnostics. | NVIDIA/ | 3.5k | — | ~3.7k | Automated safety check: Pass | Apache-2.0 | today |
| 24 | Create Earth2Studio prognostic (time-stepping forecast) model wrappers. | NVIDIA/ | 3.5k | — | ~2.6k | Automated safety check: Pass | Apache-2.0 | today |
| 25 | Maintains existing pytest and Django test suites without weakening correctness. | PostHog/ | 721 | — | ~2.5k | Automated safety check: Pass | MIT | today |
| 26 | Teaches how to write and run evals on the products/posthogai/evalharness/ harness — sandboxed agent suites that execute the real coding agent in a Docker or Modal sandbox against a seeded Hedgebox… | PostHog/ | 721 | — | ~4k | Automated safety check: Notes | MIT | today |
| 27 | A skill your agent uses when asked to reproduce a bug, verify a nightly CI failure, or confirm a failure still exists on latest source. | intel/ | 115 | — | ~7.7k | Automated safety check: Pass | Apache-2.0 | today |
| 28 | Guides an agent through reproducing, root-causing, fixing, and validating flaky tests in the PostHog monorepo. | PostHog/ | 721 | — | ~5.9k | Automated safety check: Pass | MIT | today |
| 29 | Gates whether a new test should exist and forces it to be efficient, protecting CI from low-value test bloat. | PostHog/ | 721 | — | ~5.8k | Automated safety check: Pass | MIT | today |
Community
| # | Skill | Repository | Stars | Used in | Tokens | Auto-check | Licence | Updated |
|---|---|---|---|---|---|---|---|---|
| 30 | Run Kedro's local lint / format / type-check / tests on changed files (uses the project's pre-commit hooks, ruff, mypy, pytest, lint-imports, detect-secrets, Make targets — in the right venv), or… | kedro-org/ | 11k | — | ~4k | Automated safety check: Pass | Unknown | yesterday |
| 31 | Walks through a Skyvern open-source release bump: update the version, rebuild the Python and TypeScript SDKs with Fern, commit, and open a pull request. | Skyvern-AI/ | 23k | — | ~1k | Automated safety check: Notes | AGPL-3.0 | today |
| 32 | Rules for writing, fixing and reviewing pytest unit tests that are hermetic: behavior-focused, deterministic, isolated and cheap to run. | dimensionalOS/ | 4.6k | — | ~1.4k | Automated safety check: Pass | Unknown | today |
| 33 | 33.Test Guard Reviews newly written or edited tests against nine rules that cut test bloat, such as mock-heavy checks and near-duplicate cases, before they are committed. | amElnagdy/ | 1.3k | 2 repos | ~2.1k | Automated safety check: Pass | MIT | 3 mo ago |
| 34 | Reviews an Ansible pull request by number, following the steps in the project's CLAUDE.md, with early checks for changelog fragments and tests. | ansible/ | 71k | — | ~570 | Automated safety check: Pass | GPL-3.0 | today |
| 35 | Disciplined, reproducible loop for making an FLA kernel faster (Triton, Gluon, TileLang, CuTe) without ever breaking or gaming correctness. | fla-org/ | 5.8k | — | ~2.6k | Automated safety check: Pass | MIT | yesterday |
| 36 | Creates Python projects with proper structure, virtual environments, and dependency management. | haddock-development/ | 379 | — | ~1.4k | Automated safety check: Notes | No licence | 8 mo ago |
| 37 | Sets up an isolated per-worktree Python environment for attention-gym development using nightly PyTorch and the CI-mirroring uv flow. | meta-pytorch/ | 1.3k | — | ~858 | Automated safety check: Pass | BSD-3-Clause | yesterday |
| 38 | Iterative code refinement through plan → code → evaluate → refine cycles. | EvoScientist/ | 474 | 3 repos | ~2.5k | Automated safety check: Pass | Apache-2.0 | 6 days ago |
| 39 | Run pytest tests with automatic virtual environment activation. Use this skill whenever running tests, executing pytest, or when asked to "run… | saleor/ | 23k | — | ~251 | Automated safety check: Pass | BSD-3-Clause | today |
| 40 | Developer guide for the Cortex-M (CMSIS-NN) backend in ExecuTorch: quantization pipeline, pass manager, tests and adding new ops. | pytorch/ | 5.1k | — | ~872 | Automated safety check: Pass | Unknown | today |
| 41 | Port a Node-RED node into EdgeLinkd the way this repo does it: implement the node in Rust under crates/core/src/runtime/nodes, mirror Node-RED's mocha spec as pytest tests under tests/, register the… | oldrev/ | 121 | — | ~3k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 42 | Designs and implements testing strategies for any codebase. An agent skill from CloudAI-X/claude-workflow-v2. | CloudAI-X/ | 1.4k | 1 repo | ~1.5k | Automated safety check: Pass | MIT | yesterday |
| 43 | Start, selectively modernize, fully migrate, or update Python projects using simple-modern-uv practices: uv, ruff, BasedPyright, pytest, GitHub Actions CI, and tag-driven PyPI publishing. | jlevy/ | 301 | — | ~1.9k | Automated safety check: Pass | MIT | 1 mo ago |
| 44 | Audit, plan, write, and verify unit tests for Python projects using pytest. | areed1192/ | 149 | — | ~2.6k | Automated safety check: Pass | MIT | 5 mo ago |
| 45 | 45.Chart Tests A skill your agent uses when writing, editing, reviewing, or running Helm chart tests for the Astronomer APC repository. | astronomer/ | 491 | — | ~3.2k | Automated safety check: Pass | Unknown | today |
| 46 | Guides an agent through probing an unfamiliar project's test setup, then choosing a red-light-first testing strategy matched to the task type. | huiliyi37/ | 1.1k | — | ~1k | Automated safety check: Notes | Apache-2.0 | yesterday |
| 47 | Guide for adding unit tests to AReaL. An agent skill from areal-project/AReaL. | areal-project/ | 5.8k | — | ~1.8k | Automated safety check: Pass | Apache-2.0 | 4 days ago |
| 48 | Scaffolds FastAPI projects with a layered app layout, dependency injection through Depends, async handlers and database access, middleware and pytest setup. | wshobson/ | 40k | 11 repos | ~901 | Automated safety check: Pass | MIT | 2 days ago |
Questions, answered from the data.
What is the best pytest skill?
Adk Verify Snippets (official) from google/adk-python ranks first of the 268 pytest skills listed here, with the highest score: its repository has 22k GitHub stars, its SKILL.md loads about 1.4k tokens and it passes the automated safety check with no findings. Next come Adk Setup and ONNX Runtime Test Runner.
Is there an official pytest skill?
29 of the 268 pytest skills are official, published by the vendor's own GitHub organization: Adk Verify Snippets, Adk Setup, ONNX Runtime Test Runner, Adk Agent Builder, ONNX Runtime GPU Transformers Tests and 24 more.
How are these skills ranked?
By Skill Navigator score, which combines the GitHub stars of the skill's repository (shared across that repo's skills and discounted for large collections), how many other GitHub owners carry a copy of the skill, and automated SKILL.md quality checks, minus penalties for safety-check warnings and for each further skill from the same repository. Skills that fail the safety check are not listed.