Search
pytest · Failing and flaky tests
Skills
Sort:BestMost starsTrending todayTrending this weekTrending this monthNewestRecently updatedName
| # | Skill | Repository | Stars | Used in | Tokens | Auto-check | Licence | Updated |
|---|---|---|---|---|---|---|---|---|
| 1 | Runs and debugs ONNX Runtime tests: Google Test executables for C++ and unittest or pytest for Python, with filters and build-directory guidance. | microsoft/ | 22k | — | ~1.8k | Automated safety check: Pass | MIT | today |
| 2 | Designs and implements testing strategies for any codebase. An agent skill from CloudAI-X/claude-workflow-v2. | CloudAI-X/ | 1.4k | 1 repo | ~1.5k | Automated safety check: Pass | MIT | 5 days ago |
| 3 | Guides an agent through probing an unfamiliar project's test setup, then choosing a red-light-first testing strategy matched to the task type. | huiliyi37/ | 1.1k | — | ~1k | Automated safety check: Notes | Apache-2.0 | today |
| 4 | Runs the ONNX Runtime transformers Python tests against a GPU wheel and proves the cuDNN flash attention path was used rather than a silent fallback. | microsoft/ | 22k | — | ~2.9k | Automated safety check: Pass | MIT | today |
| 5 | Run the gated test classes that plain pytest skips (slow Docker/sandbox tests, live model-provider API tests, flaky tests, trio variants). | UKGovernmentBEIS/ | 3k | — | ~1.4k | Automated safety check: Pass | MIT | today |
| 6 | A skill your agent uses when writing, editing, reviewing, or running functional (end-to-end) tests for the Astronomer APC repository. | astronomer/ | 491 | — | ~2.2k | Automated safety check: Pass | Unknown | 2 days ago |
| 7 | A skill your agent uses when selecting, running, or fixing WorldForge validation: pytest, coverage, ruff, generated provider docs, MkDocs strict build, package contract, CI failures, and release… | AbdelStark/ | 108 | — | ~871 | Automated safety check: Pass | MIT | 22 days ago |
| 8 | Test Temporal workflows with pytest, time-skipping, and mocking strategies. | wshobson/ | 40k | 12 repos | ~1.2k | Automated safety check: Pass | MIT | 6 days ago |
| 9 | A skill your agent uses when the user has concrete failing cases in code or a guardrail/classifier/filter/prompt/API they own — a red-team failure catalogue OR a CI/CD test-failure report (failing… | gaasher/ | 174 | — | ~3.6k | Automated safety check: Pass | MIT | 3 mo ago |
| 10 | Flaky test fixes: detect nondeterministic tests empirically (repeat/shuffle/parallel runs), diagnose the root cause, fix it — never retry/skip/sleep — and verify across many randomized runs. | maddhruv/ | 219 | 1 repo | ~1.2k | Automated safety check: Pass | MIT | 3 mo ago |
| 11 | Write and evaluate effective Python tests using pytest. An agent skill from iusztinpaul/squid. | iusztinpaul/ | 203 | — | ~1.3k | Automated safety check: Pass | Apache-2.0 | 1 mo ago |
| 12 | 12.Map Debug Structured MAP debugging via decomposer, actor, and monitor agents. | azalio/ | 156 | — | ~4.6k | Automated safety check: Pass | MIT | 4 days ago |
| 13 | 13.Map Debug Structured MAP debugging via task-decomposer, actor, and monitor agents. | azalio/ | 156 | — | ~4.6k | Automated safety check: Pass | MIT | 4 days ago |
| 14 | The complete QA skill for Claude Code — turn Claude into an expert QA engineer that picks the right test type, writes reliable Playwright, Cypress, and pytest tests, eliminates flaky tests, enforces… | PramodDutta/ | 235 | — | ~2.3k | Automated safety check: Pass | MIT | 7 days ago |
| 15 | Maintains existing pytest and Django test suites without weakening correctness. | PostHog/ | 40k | — | ~2.5k | Automated safety check: Pass | Unknown | yesterday |
| 16 | Guides an agent through reproducing, root-causing, fixing, and validating flaky tests in the PostHog monorepo. | PostHog/ | 40k | — | ~5.9k | Automated safety check: Pass | Unknown | yesterday |
| 17 | Exact per-package pytest/vitest/playwright commands, the registered pytest markers and which need Docker, the testutils Docker fixture matrix, the shared-test-DB isolation model and its failure… | Edwardvaneechoud/ | 385 | — | ~8.3k | Automated safety check: Notes | MIT | yesterday |
| 18 | Gates whether a new test should exist and forces it to be efficient, protecting CI from low-value test bloat. | PostHog/ | 40k | — | ~5.9k | Automated safety check: Pass | Unknown | yesterday |
| 19 | 19.Testing A skill your agent uses when writing, fixing, reviewing, or planning tests in any project, in whatever runner the project already uses. | WrongStack/ | 371 | — | ~2.3k | Automated safety check: Pass | MIT | yesterday |
| 20 | Testing conventions, patterns, and automation for Agent Kernel development. | yaalalabs/ | 192 | — | ~15k | Automated safety check: Pass | Apache-2.0 | 2 days ago |
| 21 | 21.Write Tests Write or revise tests with an emphasis on behavior, regression coverage, pytest style, and avoiding "slop tests." Use when adding tests, fixing failing tests, reviewing test quality, or deciding… | open-thoughts/ | 301 | — | ~498 | Automated safety check: Pass | Apache-2.0 | 12 days ago |
| 22 | Identifies non-deterministic or unreliable tests through static code analysis and test result analysis. | ArabelaTso/ | 253 | — | ~2k | Automated safety check: Pass | Apache-2.0 | 1 mo ago |
| 23 | Write and evaluate effective Python tests using pytest. An agent skill from benchflow-ai/skillsbench. | benchflow-ai/ | 1.8k | — | ~1.3k | Automated safety check: Pass | Apache-2.0 | 2 mo ago |
| 24 | Testing strategy: pyramid, AAA, mocks/fakes/stubs, flaky tests, coverage. | softspark/ | 179 | — | ~1.6k | Automated safety check: Pass | Apache-2.0 | 3 days ago |
| 25 | A skill your agent uses when asked to reproduce a bug, verify a nightly CI failure, or confirm a failure still exists on latest source. | intel/ | 115 | — | ~7.7k | Automated safety check: Pass | Apache-2.0 | yesterday |