Topic · Testing & QA
Best failing and flaky tests skills, page 11
Failing and flaky tests skills, ranked
Ranked by score. Sort bymost stars,trending,newest,recently updated
| # | Skill | Repository | Stars | Used in | Tokens | Auto-check | Licence | Updated |
|---|---|---|---|---|---|---|---|---|
| 481 | Automatically performs git bisect to identify the first bad commit that introduced a bug or failure. | ArabelaTso/ | 253 | — | ~1.2k | Automated safety check: Pass | Apache-2.0 | 1 mo ago |
| 482 | 482.E2E Testing Playwright E2E 测试模式、页面对象模型(POM)、配置、CI/CD 集成、产物管理以及不稳定测试(flaky test)策略。 | xu-xiang/ | 2k | — | ~1.9k | Automated safety check: Pass | MIT | 7 mo ago |
| 483 | 483.E2E Testing Playwright E2E 测试模式、页面对象模型 (Page Object Model)、配置、CI/CD 集成、产物管理以及不稳定测试 (Flaky Test) 策略。 | xu-xiang/ | 2k | — | ~2k | Automated safety check: Pass | MIT | 7 mo ago |
| 484 | 484.Error Analysis Evals-first error analysis for LLM apps: clusters real Langfuse or JSONL traces into a human-confirmed failure taxonomy with counts, then recommends binary pass/fail evals for recurring named modes. | yonatangross/ | 292 | — | ~3.6k | Automated safety check: Notes | MIT | today |
| 485 | Deep technical debt and CI stability audit for identifying test theater, missing or mis-scoped tests, actual and potential test failures, flaky-test risk, dependency/toolchain brittleness, and… | ZaxbyHub/ | 494 | — | ~2.6k | Automated safety check: Pass | MIT | today |
| 486 | Systematic methodology for debugging bugs, test failures, and unexpected behavior. | aiskillstore/ | 433 | — | ~1.4k | Automated safety check: Pass | No licence | today |
| 487 | Root cause analysis for debugging. An agent skill from aiskillstore/marketplace. | aiskillstore/ | 433 | — | ~1.1k | Automated safety check: Pass | No licence | today |
| 488 | Triage and fix PR CI failures caused by the PR's own changes. | openshift-eng/ | 120 | — | ~3.1k | Automated safety check: Pass | Apache-2.0 | 3 days ago |
| 489 | Check CI status on a GitHub PR and detect new failures. An agent skill from openshift-eng/ai-helpers. | openshift-eng/ | 120 | — | ~885 | Automated safety check: Pass | Apache-2.0 | 3 days ago |
| 490 | 490.Manor CI Triage A skill your agent uses when Manor GitHub Actions, .github/workflows/ci.yml, OSS smoke/regression jobs, web source smoke, frontend build, lint, or public CI failure logs need diagnosis or repair. | manor-os/ | 162 | — | ~533 | Automated safety check: Pass | Unknown | 1 mo ago |
| 491 | 491.Test Gradle Run tests headless and return only failing test output. An agent skill from jvm-skills/jvm-skills. | jvm-skills/ | 139 | — | ~314 | Automated safety check: Pass | Apache-2.0 | 1 mo ago |
| 492 | 492.Reproduce First Reproduce first, then fix — turn a reported defect into a failing test, fix it, and confirm that same test goes green. | vfarcic/ | 109 | — | ~2k | Automated safety check: Pass | MIT | today |
| 493 | 493.Actions Manager GitHub Actions: workflow runs, logs, re-runs and CI failure triage. | Community-Access/ | 423 | — | ~904 | Automated safety check: Pass | MIT | 16 days ago |
| 494 | Investigate a failing integration test from a GitHub issue. An agent skill from microsoft/GitHub-Copilot-for-Azure. | microsoft/ | 255 | — | ~295 | Automated safety check: Pass | MIT | today |
| 495 | Detect flaky test detector operations. An agent skill from jeremylongshore/tons-of-skills-marketplace. | jeremylongshore/ | 2.8k | — | ~554 | Automated safety check: Pass | MIT | today |
| 496 | Finds the root cause of a bug, test failure, or unexpected behavior before proposing any fix. | cbrock84/ | 2k | — | ~657 | Automated safety check: Pass | MIT | 22 days ago |
| 497 | 497.Build Fix Workflow for incrementally fixing TypeScript and build errors. | jh941213/ | 125 | — | ~326 | Automated safety check: Notes | No licence | 2 mo ago |
| 498 | 498.Debug Loop A skill your agent uses when debugging a failing test or runtime error with hypothesis-driven investigation, autonomous command validation, and systematic root cause elimination. | proffesor-for-testing/ | 495 | — | ~545 | Automated safety check: Pass | MIT | today |
| 499 | 499.Qcsd Cicd Swarm A skill your agent uses when enforcing CI/CD quality gates before release, running regression analysis, detecting flaky tests, or assessing deployment readiness in the QCSD Verification phase. | proffesor-for-testing/ | 495 | — | ~2.3k | Automated safety check: Pass | MIT | today |
| 500 | Quality Engineering iteration loops for autonomous test improvement, coverage achievement, and quality gate compliance. | proffesor-for-testing/ | 495 | — | ~3.2k | Automated safety check: Pass | MIT | today |
| 501 | Design and implement effective test automation with proper pyramid, patterns, and CI/CD integration. | proffesor-for-testing/ | 495 | — | ~1.6k | Automated safety check: Pass | MIT | today |
| 502 | Apply Edward de Bono's Six Thinking Hats methodology to software testing for comprehensive quality analysis. | proffesor-for-testing/ | 495 | — | ~1.9k | Automated safety check: Pass | MIT | today |
| 503 | A skill your agent uses when a test is failing and you need to determine root cause: is it flaky, an environment issue, or a real regression? | proffesor-for-testing/ | 495 | — | ~811 | Automated safety check: Pass | MIT | today |
| 504 | A skill your agent uses when querying test history, analyzing flakiness rates, tracking MTTR, or building quality trend dashboards from test execution data. | proffesor-for-testing/ | 495 | — | ~718 | Automated safety check: Pass | MIT | today |
| 505 | 505.Bisect Bisect regressions or broken builds in the Chromium repository using deterministic binary search, checking out midpoint commits, syncing dependencies with gclient sync, compiling the target with… | nwjs/ | 160 | — | ~1.6k | Automated safety check: Pass | BSD-3-Clause | today |
| 506 | 506.Gsd Debug A skill your agent uses when a bug, failing test, or unexpected behaviour needs investigation, especially across sessions or context resets. | coco-research/ | 513 | — | ~1.3k | Automated safety check: Notes | Unknown | today |
| 507 | Selectively instruments code to capture runtime data for debugging failures and bugs. | ArabelaTso/ | 253 | — | ~2.1k | Automated safety check: Pass | Apache-2.0 | 1 mo ago |
| 508 | Automatically generate clear, actionable issue reports from failing tests and repository analysis. | ArabelaTso/ | 253 | — | ~3.5k | Automated safety check: Pass | Apache-2.0 | 1 mo ago |
| 509 | Locate root causes of failing regression tests by analyzing code changes, error messages, and test dependencies. | ArabelaTso/ | 253 | — | ~3.4k | Automated safety check: Pass | Apache-2.0 | 1 mo ago |
| 510 | Instruments programs to record execution information for deterministic replay debugging. | ArabelaTso/ | 253 | — | ~2.5k | Automated safety check: Pass | Apache-2.0 | 1 mo ago |
| 511 | Analyze failing tests to detect functional bugs in code. An agent skill from ArabelaTso/Skills-4-SE. | ArabelaTso/ | 253 | — | ~2.8k | Automated safety check: Pass | Apache-2.0 | 1 mo ago |
| 512 | Explain why counterexamples violate specifications by analyzing formal specifications (temporal logic, invariants, pre/postconditions, code contracts), informal requirements (user stories… | ArabelaTso/ | 253 | — | ~3.5k | Automated safety check: Pass | Apache-2.0 | 1 mo ago |
| 513 | Codex adapter for deep technical-debt and CI-stability audits. | ZaxbyHub/ | 494 | — | ~239 | Automated safety check: Pass | MIT | today |
| 514 | A skill your agent uses when diagnosing a bug, failing test, build issue, or unexpected behavior in Cursor before proposing a fix. | 6BNBN/ | 134 | — | ~332 | Automated safety check: Pass | MIT | 7 mo ago |
| 515 | Diagnoses dbt Cloud/platform job failures by analyzing run logs, querying the Admin API, reviewing git history, and investigating data issues. | Kilo-Org/ | 190 | — | ~2.6k | Automated safety check: Pass | Apache-2.0 | 11 days ago |
| 516 | 516.Implement A skill your agent uses when an approved plan and task list exist and it is time to turn them into working, tested code — the SDD phase after analyze and before verify. | ericrisco/ | 180 | — | ~5.9k | Automated safety check: Notes | MIT | today |
| 517 | 517.Bug Diagnosis Diagnose a bug systematically instead of guessing — reproduce, isolate, form hypotheses, and test them to root cause. | mohitagw15856/ | 1.4k | — | ~804 | Automated safety check: Pass | MIT | today |
| 518 | 518.Regex Builder Build a regular expression from a plain-English description, or explain an existing one. | mohitagw15856/ | 1.4k | — | ~704 | Automated safety check: Pass | MIT | today |
| 519 | 519.Stop Digging Anti-thrashing circuit-breaker. An agent skill from ccplugins/awesome-claude-code-plugins. | ccplugins/ | 970 | — | ~1.5k | Automated safety check: Pass | MIT | 1 mo ago |
| 520 | 520.Revision Trace bugs and manual fixes back to kits and prompts, then fix at the source so the iteration loop can reproduce the fix autonomously. | hashgraph-online/ | 1.3k | — | ~4.9k | Automated safety check: Pass | Apache-2.0 | today |
| 521 | When multiple tests fail, assign each failing test file to a separate subagent that fixes it independently in parallel. | spencerpauly/ | 844 | — | ~525 | Automated safety check: Pass | CC0-1.0 | 2 mo ago |
| 522 | Testing patterns for PHPUnit and Playwright E2E tests. An agent skill from Microck/ordinary-claude-skills. | Microck/ | 404 | — | ~712 | Automated safety check: Pass | Unknown | 1 mo ago |
| 523 | 523.Testing Run and troubleshoot tests for DBHub, including unit tests, integration tests with Testcontainers, and database-specific tests. | Microck/ | 404 | — | ~854 | Automated safety check: Pass | Unknown | 1 mo ago |
| 524 | Apply systems thinking — causal loop diagrams, stock-and-flow models, system archetypes, and leverage-point analysis — to organizational, economic, or social problems where feedback loops, delays… | asgard-ai-platform/ | 242 | — | ~1.2k | Automated safety check: Pass | MIT | 4 mo ago |
| 525 | Multi-agent Test-Driven Development workflow. An agent skill from nwjs/chromium.src. | nwjs/ | 160 | — | ~1.4k | Automated safety check: Pass | BSD-3-Clause | today |
| 526 | 526.Testing Writing effective tests and running them successfully. An agent skill from rsmdt/the-startup. | rsmdt/ | 560 | — | ~1.2k | Automated safety check: Pass | MIT | 2 mo ago |
| 527 | 527.Validate Fix Prove a fix works before declaring done — re-run the failing test, run the full suite, typecheck, lint, and harden against recurrence. | danielvm-git/ | 260 | — | ~1.1k | Automated safety check: Pass | MIT | 18 days ago |
| 528 | 528.CI Debug Diagnose a failing CI run against an 11-pattern playbook. An agent skill from yonatangross/orchestkit. | yonatangross/ | 292 | — | ~2.7k | Automated safety check: Notes | MIT | today |