Topic · Testing & QA

Best failing and flaky tests skills, page 10

Skills #433–480 of 554, ranked by score.

Failing and flaky tests skills, ranked

Ranked by score. Sort bymost stars,trending,newest,recently updated

Failing and flaky tests skills, ranked
#SkillRepositoryStarsUsed inTokensAuto-checkLicenceUpdated
433

Apply when extracting code from a large monolith file into submodules.

ZaxbyHub/opencode-swarm493—~2.6kAutomated safety check: PassMITtoday
434

End-to-end CI monitor that takes an already-human-reviewed PR, exhaustively researches every CI failure, fixes it end-to-end, iterates until all required checks are green (max 5 fix cycles), then…

ZaxbyHub/opencode-swarm493—~4.9kAutomated safety check: PassMITtoday
435

Content-assertion sweep after editing SKILL.md files. An agent skill from ZaxbyHub/opencode-swarm.

ZaxbyHub/opencode-swarm493—~532Automated safety check: PassMITtoday
436

Write or revise tests with an emphasis on behavior, regression coverage, pytest style, and avoiding "slop tests." Use when adding tests, fixing failing tests, reviewing test quality, or deciding…

open-thoughts/OpenThoughts-Agent301—~498Automated safety check: PassApache-2.011 days ago
437
437.Ut CheckOfficial

Analyze UT (unit test) results for a torch-xpu-ops PR. An agent skill from intel/torch-xpu-ops.

intel/torch-xpu-ops115—~1.6kAutomated safety check: PassApache-2.0today
438

A skill your agent uses when an agent says a test, VM, proof, evaluator, or release gate is overloaded, too strict, blocking staging, or causing false positives; split checks by profile, measure the…

AnastasiyaW/codex-claude-code-config154—~1kAutomated safety check: PassMITyesterday
439

Debug and fix a bug by starting from one concrete failure, forming a falsifiable local hypothesis, making the smallest grounded edit, and immediately running focused validation.

lge-ros2/cloisim176—~1.2kAutomated safety check: PassMIT4 days ago
440

Inspect GitHub PR for CI failures, merge conflicts, update-branch requirements, reviewer comments, change requests, and unresolved review threads.

akiojin/unity-cli107—~4.6kAutomated safety check: PassApache-2.04 days ago
441

Weekly. An agent skill from markfulton/ai-employees.

markfulton/ai-employees498—~14kAutomated safety check: PassMIT2 days ago
442

Analyzes DevOps Center test failures and Code Analyzer violations in plain language — failure category, offending file/class/method/line, rule violated, fix direction, and prioritized improvement…

forcedotcom/sf-skills1.1k—~1.4kAutomated safety check: PassApache-2.02 days ago
443

Orchestrates test suite execution with parallel sharding, intelligent retry, and real-time reporting across Jest, Vitest, and Playwright.

proffesor-for-testing/agentic-qe495—~1.2kAutomated safety check: PassMIT5 days ago
444

Design and implement effective test automation with proper pyramid, patterns, and CI/CD integration.

proffesor-for-testing/agentic-qe495—~1.4kAutomated safety check: PassMIT5 days ago
445

Guide for disabling Chromium tests (C++ or WebUI). An agent skill from nwjs/chromium.src.

nwjs/chromium.src160—~790Automated safety check: PassBSD-3-Clause6 days ago
446

Identifies non-deterministic or unreliable tests through static code analysis and test result analysis.

ArabelaTso/Skills-4-SE253—~2kAutomated safety check: PassApache-2.01 mo ago
447

Automatically updates a codebase to a new language version, framework version, or library update while ensuring all tests still pass.

ArabelaTso/Skills-4-SE253—~1.8kAutomated safety check: PassApache-2.01 mo ago
448

Write and evaluate effective Python tests using pytest. An agent skill from benchflow-ai/skillsbench.

benchflow-ai/skillsbench1.8k—~1.3kAutomated safety check: PassApache-2.02 mo ago
449

Comprehensive testing guide for Cloudflare Workers using Vitest and @cloudflare/vitest-pool-workers.

secondsky/claude-skills227—~5.3kAutomated safety check: PassMIT11 days ago
450

Test-first development route for TDD, writing failing tests first, RED - GREEN - REFACTOR, and behavior-changing feature/bug/refactor work.

foryourhealth111-pixel/Vibe-Skills3.6k—~455Automated safety check: PassApache-2.01 mo ago
451

Testing strategy: pyramid, AAA, mocks/fakes/stubs, flaky tests, coverage.

softspark/ai-toolkit179—~1.6kAutomated safety check: PassApache-2.0yesterday
452
452.Trace

A skill your agent uses when encountering bugs, test failures, runtime errors, unexpected behavior, broken builds, or "this doesn't work" reports.

ccplugins/awesome-claude-code-plugins970—~1.4kAutomated safety check: PassApache-2.01 mo ago
453

Write or repair Elixir tests with ExUnit, sandbox isolation, async reliability, Mox, ExMachina, and LiveViewTest.

oliver-kriska/claude-elixir-phoenix565—~750Automated safety check: PassMIT4 days ago
454

Minimum-requirements checklist for any change — code or docs-only.

oocx/tfplan2md174—~1.8kAutomated safety check: PassMIT2 days ago
455

Red-green-refactor scaffold for building new features with TDD.

gustavscirulis/snapgrid1171 repo~2.4kAutomated safety check: NotesUnknown5 mo ago
456

Isolate complex Adobe failures across DNS/TLS, auth, entitlement, schema, async state, storage, Runtime, and downstream layers without leaking credentials or content.

jeremylongshore/tons-of-skills-marketplace2.8k—~1.1kAutomated safety check: PassMITyesterday
457

Isolate difficult Canva Connect failures without rotating credentials, replaying writes, or leaking customer data.

jeremylongshore/tons-of-skills-marketplace2.8k—~1kAutomated safety check: PassMITyesterday
458

Deep debugging for Figma API issues: network analysis, response inspection, and support escalation.

jeremylongshore/tons-of-skills-marketplace2.8k—~1.9kAutomated safety check: PassMITyesterday
459

Debug complex Shopify API issues using cost analysis, request tracing, webhook delivery inspection, and GraphQL introspection.

jeremylongshore/tons-of-skills-marketplace2.8k—~1.3kAutomated safety check: PassMITyesterday
460
460.Trace

A skill your agent uses when encountering bugs, test failures, runtime errors, broken builds, or "this doesn't work" reports.

jeremylongshore/tons-of-skills-marketplace2.8k—~5kAutomated safety check: PassMITyesterday
461

Advanced debugging for hard-to-diagnose Vercel issues including cold starts, edge errors, and function tracing.

jeremylongshore/tons-of-skills-marketplace2.8k—~2.1kAutomated safety check: PassMITyesterday
462

Monitor a PR and fix CI or review feedback. An agent skill from BuilderIO/agent-native.

BuilderIO/agent-native7.1k—~8.5kAutomated safety check: PassNo licencetoday
463

End-to-end testing with Playwright: test generation, page objects, locator strategy, flaky- test diagnosis, visual regression, and CI integration.

borghei/Claude-Skills886—~2kAutomated safety check: PassMIT2 days ago
464

Diagnose a failing CI workflow run (lint, markdown lint, build, or unit tests) — identify which job failed, the cause, and a concrete fix

adamayoung/TMDb178—~1.6kAutomated safety check: PassApache-2.06 days ago
465

Diagnose a PR, branch, run, or GitHub Actions URL from CI evidence and produce a fix plan without changing code.

Terry-Mao/AICodingFlow167—~447Automated safety check: PassMIT6 days ago
466
466.Openai Gh Fix CIOfficial

A skill your agent uses when a user asks to debug or fix failing GitHub PR checks that run in GitHub Actions; use gh to inspect checks and logs, summarize failure context, draft a fix plan, and…

trailofbits/skills-curated513—~955Automated safety check: NotesCC-BY-SA-4.02 mo ago
467

Running and debugging tests in the Golem workspace. An agent skill from golemcloud/golem.

golemcloud/golem1.5k—~1.8kAutomated safety check: PassUnknowntoday
468

Codex adapter for monitoring and fixing CI failures on opencode-swarm PRs.

ZaxbyHub/opencode-swarm493—~551Automated safety check: PassMITtoday
469

A skill your agent uses when asked to trace, investigate, root-cause, reproduce, plan, fix, resolve, close, or prepare a PR for an issue, bug report, defect, regression, failing test, or confusing…

ZaxbyHub/opencode-swarm493—~383Automated safety check: PassMITtoday
470

A skill your agent uses when asked to trace, investigate, root-cause, reproduce, plan, fix, resolve, close, or prepare a PR for an issue, bug report, defect, regression, failing test, or confusing…

ZaxbyHub/opencode-swarm493—~397Automated safety check: PassMITtoday
471

run failing tests immediately after testgate failure. An agent skill from ZaxbyHub/opencode-swarm.

ZaxbyHub/opencode-swarm493—~332Automated safety check: PassMITtoday
472

Iterate on a GitHub pull request — drive it through CI, code review, and QA until it is merge-ready.

OpenHands/extensions161—~3.8kAutomated safety check: PassMITyesterday
473
473.Fix

Fix bugs and broken behavior when there is enough evidence to act on a repair path.

avibebuilder/claude-prime120—~1kAutomated safety check: PassMIT4 mo ago
474

Deliver changes in this monorepo through Git-Spice, focused checks, Woodpecker CI, automated review, and PR evidence.

shepherdjerred/monorepo112—~875Automated safety check: PassGPL-3.0today
475
475.Fix ReproduceOfficial

A skill your agent uses when asked to reproduce a bug, verify a nightly CI failure, or confirm a failure still exists on latest source.

intel/torch-xpu-ops115—~7.7kAutomated safety check: PassApache-2.0today
476

A skill your agent uses whenever the design will encounter conditions that can't be fully predicted at design time — load spikes, edge-case content, real-world variability, hardware variation…

hashgraph-online/awesome-codex-plugins1.3k—~3.5kAutomated safety check: PassApache-2.0yesterday
477

Load when investigating a failing PR CI pipeline or checking PR health.

datadog-labs/agent-skills177—~2.7kAutomated safety check: PassMITyesterday
478

Systematically fix all failing tests after business logic changes or refactoring

NeoLabHQ/context-engineering-kit1.7k—~914Automated safety check: PassGPL-3.01 mo ago
479

Production incidents dashboard. An agent skill from davepoon/buildwithclaude.

davepoon/buildwithclaude3.6k—~1.8kAutomated safety check: NotesMITtoday
480

Autonomous PR merge pipeline. An agent skill from davepoon/buildwithclaude.

davepoon/buildwithclaude3.6k—~3.1kAutomated safety check: NotesMITtoday