Search

pytest · Failing and flaky tests

25 skills found.
Search results
#SkillRepositoryStarsUsed inTokensAuto-checkLicenceUpdated
1

Runs and debugs ONNX Runtime tests: Google Test executables for C++ and unittest or pytest for Python, with filters and build-directory guidance.

microsoft/onnxruntime22k—~1.8kAutomated safety check: PassMITtoday
2

Designs and implements testing strategies for any codebase. An agent skill from CloudAI-X/claude-workflow-v2.

CloudAI-X/claude-workflow-v21.4k1 repo~1.5kAutomated safety check: PassMIT5 days ago
3

Guides an agent through probing an unfamiliar project's test setup, then choosing a red-light-first testing strategy matched to the task type.

huiliyi37/Tianshu-harness1.1k—~1kAutomated safety check: NotesApache-2.0today
4

Runs the ONNX Runtime transformers Python tests against a GPU wheel and proves the cuDNN flash attention path was used rather than a silent fallback.

microsoft/onnxruntime22k—~2.9kAutomated safety check: PassMITtoday
5

Run the gated test classes that plain pytest skips (slow Docker/sandbox tests, live model-provider API tests, flaky tests, trio variants).

UKGovernmentBEIS/inspect_ai3k—~1.4kAutomated safety check: PassMITtoday
6

A skill your agent uses when writing, editing, reviewing, or running functional (end-to-end) tests for the Astronomer APC repository.

astronomer/astronomer491—~2.2kAutomated safety check: PassUnknown2 days ago
7

A skill your agent uses when selecting, running, or fixing WorldForge validation: pytest, coverage, ruff, generated provider docs, MkDocs strict build, package contract, CI failures, and release…

AbdelStark/worldforge108—~871Automated safety check: PassMIT22 days ago
8

Test Temporal workflows with pytest, time-skipping, and mocking strategies.

wshobson/agents40k12 repos~1.2kAutomated safety check: PassMIT6 days ago
9

A skill your agent uses when the user has concrete failing cases in code or a guardrail/classifier/filter/prompt/API they own — a red-team failure catalogue OR a CI/CD test-failure report (failing…

gaasher/Agent-Loop-Skills174—~3.6kAutomated safety check: PassMIT3 mo ago
10

Flaky test fixes: detect nondeterministic tests empirically (repeat/shuffle/parallel runs), diagnose the root cause, fix it — never retry/skip/sleep — and verify across many randomized runs.

maddhruv/absolute2191 repo~1.2kAutomated safety check: PassMIT3 mo ago
11

Write and evaluate effective Python tests using pytest. An agent skill from iusztinpaul/squid.

iusztinpaul/squid203—~1.3kAutomated safety check: PassApache-2.01 mo ago
12

Structured MAP debugging via decomposer, actor, and monitor agents.

azalio/map-framework156—~4.6kAutomated safety check: PassMIT4 days ago
13

Structured MAP debugging via task-decomposer, actor, and monitor agents.

azalio/map-framework156—~4.6kAutomated safety check: PassMIT4 days ago
14

The complete QA skill for Claude Code — turn Claude into an expert QA engineer that picks the right test type, writes reliable Playwright, Cypress, and pytest tests, eliminates flaky tests, enforces…

PramodDutta/qaskills235—~2.3kAutomated safety check: PassMIT7 days ago
15

Maintains existing pytest and Django test suites without weakening correctness.

PostHog/posthog40k—~2.5kAutomated safety check: PassUnknownyesterday
16

Guides an agent through reproducing, root-causing, fixing, and validating flaky tests in the PostHog monorepo.

PostHog/posthog40k—~5.9kAutomated safety check: PassUnknownyesterday
17

Exact per-package pytest/vitest/playwright commands, the registered pytest markers and which need Docker, the testutils Docker fixture matrix, the shared-test-DB isolation model and its failure…

Edwardvaneechoud/Flowfile385—~8.3kAutomated safety check: NotesMITyesterday
18
18.Writing TestsOfficial

Gates whether a new test should exist and forces it to be efficient, protecting CI from low-value test bloat.

PostHog/posthog40k—~5.9kAutomated safety check: PassUnknownyesterday
19

A skill your agent uses when writing, fixing, reviewing, or planning tests in any project, in whatever runner the project already uses.

WrongStack/WrongStack371—~2.3kAutomated safety check: PassMITyesterday
20

Testing conventions, patterns, and automation for Agent Kernel development.

yaalalabs/agent-kernel192—~15kAutomated safety check: PassApache-2.02 days ago
21

Write or revise tests with an emphasis on behavior, regression coverage, pytest style, and avoiding "slop tests." Use when adding tests, fixing failing tests, reviewing test quality, or deciding…

open-thoughts/OpenThoughts-Agent301—~498Automated safety check: PassApache-2.012 days ago
22

Identifies non-deterministic or unreliable tests through static code analysis and test result analysis.

ArabelaTso/Skills-4-SE253—~2kAutomated safety check: PassApache-2.01 mo ago
23

Write and evaluate effective Python tests using pytest. An agent skill from benchflow-ai/skillsbench.

benchflow-ai/skillsbench1.8k—~1.3kAutomated safety check: PassApache-2.02 mo ago
24

Testing strategy: pyramid, AAA, mocks/fakes/stubs, flaky tests, coverage.

softspark/ai-toolkit179—~1.6kAutomated safety check: PassApache-2.03 days ago
25
25.Fix ReproduceOfficial

A skill your agent uses when asked to reproduce a bug, verify a nightly CI failure, or confirm a failure still exists on latest source.

intel/torch-xpu-ops115—~7.7kAutomated safety check: PassApache-2.0yesterday