Search
Development · Failing and flaky tests
Skills
Sort:BestMost starsTrending todayTrending this weekTrending this monthNewestRecently updatedName
| # | Skill | Repository | Stars | Used in | Tokens | Auto-check | Licence | Updated |
|---|---|---|---|---|---|---|---|---|
| 1 | Watches an open GitHub pull request until it merges, handling review comments, diagnosing CI failures and retrying flaky checks along the way. | openinterpreter/ | 69k | 3 repos | ~4.2k | Automated safety check: Pass | Apache-2.0 | today |
| 2 | Investigates failing Pester tests in PowerShell CI jobs by following a six-step workflow from pull request status to documented fix recommendations. | PowerShell/ | 56k | — | ~5.1k | Automated safety check: Pass | MIT | yesterday |
| 3 | A skill your agent uses when encountering any bug, test failure, or unexpected behavior, before proposing fixes | ultralisp/ | 258 | 52 repos | ~2.4k | Automated safety check: Pass | No licence | 27 days ago |
| 4 | Runs RustPython tests inside a Linux container built with Apple's container CLI, so macOS users can compare Linux results with their local ones. | RustPython/ | 22k | — | ~467 | Automated safety check: Pass | MIT | today |
| 5 | Iterate on a PR until CI passes. An agent skill from meshery/meshery-operator. | meshery/ | 151 | 7 repos | ~2.2k | Automated safety check: Pass | Apache-2.0 | 20 days ago |
| 6 | Takes a GitHub or YouTrack issue for the Exposed project through reproduction, a failing test, a fix, validation and a pull request. | JetBrains/ | 9.3k | — | ~3.8k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 7 | Handle a Perfherder performance regression bug end to end: read the alert bug, confirm whether the regression is real, find the cause, and iterate to a fix. | mozilla-firefox/ | 13k | — | ~1.7k | Automated safety check: Pass | Unknown | today |
| 8 | 8.Testing A skill your agent uses for every Kortix test task, behavior change, bug fix, refactor, API route change, CLI change, SDK change, browser journey, test failure, coverage question, local benchmark… | kortix-ai/ | 20k | — | ~3.6k | Automated safety check: Notes | Unknown | today |
| 9 | Applies a four-phase debugging routine that finds the root cause of a bug or failing test before any fix is written. | ChrisWiles/ | 6.1k | 3 repos | ~1.2k | Automated safety check: Pass | No licence | 9 mo ago |
| 10 | Fixes a React Router bug reported in a GitHub issue end to end: fetching the issue, validating the reproduction, writing a failing test and implementing the fix on a new branch. | remix-run/ | 57k | — | ~1.3k | Automated safety check: Pass | MIT | yesterday |
| 11 | Applies a stop-the-line rule and a step-by-step triage when tests fail, builds break or something stops working, aiming at the root cause instead of guesses. | addyosmani/ | 105k | 1 repo | ~2.6k | Automated safety check: Pass | MIT | yesterday |
| 12 | Plans the smallest check that could disprove a code change in the OpenLogi project, then escalates through reproduction, focused tests and a final gate before a push. | AprilNEA/ | 23k | — | ~1.4k | Automated safety check: Pass | Apache-2.0 | today |
| 13 | Upgrades a Python standard library module from CPython into RustPython with update_lib, then triages and marks the tests that still fail. | RustPython/ | 22k | — | ~876 | Automated safety check: Pass | MIT | today |
| 14 | A skill your agent uses when encountering any bug, test failure, or unexpected behavior during spec-superflow execution, before proposing fixes. | MageByte-Zero/ | 846 | 1 repo | ~1.6k | Automated safety check: Pass | MIT | 9 days ago |
| 15 | Debug and verification workflow for runtime-bundle and module-resolution regressions. | vercel/ | 143k | 1 repo | ~618 | Automated safety check: Pass | MIT | today |
| 16 | Runs a gated finish-line checklist before committing a PlotJuggler PJ4 change: build proof, red-test triage, hooks, docs freshness and a diff self-review. | PlotJuggler/ | 6.2k | — | ~1.3k | Automated safety check: Pass | MPL-2.0 | 9 days ago |
| 17 | Makes the smallest code change that fixes one well-scoped problem, such as a CI failure, review comment or typo, without refactoring anything unrelated. | cobusgreyling/ | 11k | 1 repo | ~671 | Automated safety check: Notes | MIT | today |
| 18 | 18.PR Review Address review comments and CI failures for the current branch's PR | wysaid/ | 1.9k | — | ~1.4k | Automated safety check: Pass | MIT | 2 mo ago |
| 19 | A skill your agent uses when encountering any bug, test failure, or unexpected behavior, before proposing fixes - four-phase framework (root cause investigation, pattern analysis, hypothesis… | ed3dai/ | 250 | 3 repos | ~2.4k | Automated safety check: Pass | No licence | 1 mo ago |
| 20 | Guides systematic root-cause debugging. An agent skill from abashev/vfs-s3. | abashev/ | 106 | 6 repos | ~2.6k | Automated safety check: Pass | Apache-2.0 | 4 days ago |
| 21 | Investigates a failing RustPython test by comparing it with CPython, then either fixes it or gathers the details for an incompatibility report. | RustPython/ | 22k | — | ~467 | Automated safety check: Pass | MIT | today |
| 22 | 22.Papercuts Log genuine, recurring repository friction to .agents/PAPERCUTS.md — confusing setup, a flaky repo command or script, a misleading in-repo error, stale generated files, or a non-obvious gotcha that… | every-app/ | 23k | — | ~1.2k | Automated safety check: Pass | MIT | 2 days ago |
| 23 | After a bug is found, traces its root cause and feeds a new testable invariant back into the project spec so the bug class can't recur. | JuliusBrussee/ | 1.2k | — | ~653 | Automated safety check: Pass | MIT | 1 mo ago |
| 24 | 24.Veomni Debug A skill your agent uses for ANY bug, error, crash, wrong output, loss divergence, gradient explosion, test failure, CUDA error, distributed training hang, checkpoint load failure, or unexpected… | ByteDance-Seed/ | 2.2k | — | ~2.8k | Automated safety check: Pass | Apache-2.0 | today |
| 25 | 25.Skillhone Local Issue, pull-request, and Wiki workbench for agent skills. | Tencent/ | 169 | — | ~3.4k | Automated safety check: Pass | Unknown | 21 days ago |
| 26 | A skill your agent uses when the user asks to fix a bug, references a GitHub issue number, or describes an issue and wants a fix. | brunosabot/ | 269 | — | ~529 | Automated safety check: Pass | MIT | 4 mo ago |
| 27 | Enforce Sentry Dart/Flutter SDK test conventions for naming, structure, and fixtures. | getsentry/ | 873 | — | ~3.1k | Automated safety check: Pass | MIT | yesterday |
| 28 | 28.Reviewloop Iteratively improves a PR until all review bots (Greptile, Devin, and others) are satisfied with zero unresolved comments, then fixes any CI failures. | ankitvgupta/ | 495 | — | ~2.3k | Automated safety check: Pass | MIT | 1 mo ago |
| 29 | 29.MCP Debugger A skill your agent uses when investigating a bug, failing test, or unexpected runtime behavior and the mcp-debugger MCP server is available — drives real step-through debuggers (breakpoints, stack… | debugmcp/ | 174 | — | ~4.2k | Automated safety check: Pass | MIT | today |
| 30 | 30.Dbg Debug applications using the dbg CLI debugger. An agent skill from theodo-group/debug-that. | theodo-group/ | 158 | — | ~2.5k | Automated safety check: Pass | MIT | yesterday |
| 31 | 31.CI Triage Triage failing GitHub PR checks: list failures with gh, fetch capped Actions logs, skip non-Actions checks, and summarize root cause. | Mentra-Community/ | 2.4k | — | ~582 | Automated safety check: Pass | Apache-2.0 | today |
| 32 | 32.Releasing Version and release c15t packages with Changesets. An agent skill from c15t/c15t. | c15t/ | 1.9k | — | ~753 | Automated safety check: Pass | Apache-2.0 | yesterday |
| 33 | Guides an agent through probing an unfamiliar project's test setup, then choosing a red-light-first testing strategy matched to the task type. | huiliyi37/ | 1.1k | — | ~1k | Automated safety check: Notes | Apache-2.0 | today |
| 34 | Investigate and fix flaky/random CI test failures in dotnet/macios. | dotnet/ | 2.9k | — | ~1.3k | Automated safety check: Pass | Unknown | 2 days ago |
| 35 | 35.Testing Run tests and add Next.js version coverage for the cache handler. | trieb-work/ | 151 | — | ~1.1k | Automated safety check: Pass | MIT | 2 days ago |
| 36 | Runs a reproduce, localize, hypothesize, test, fix and verify loop to find a bug's root cause, applies the minimal fix and hands off a regression test. | jsmastery-pro/ | 1.5k | — | ~1.8k | Automated safety check: Notes | MIT | 2 mo ago |
| 37 | 37.Code Solving Structured coding workflow for non-trivial code work: debug, build features, refactor, optimize, migrate and review code through 7 steps with evidence-based quality gates. | HoangTheQuyen/ | 122 | — | ~3.7k | Automated safety check: Pass | MIT | 2 days ago |
| 38 | 38.Fix CI Run a local pnpm monorepo CI loop, fix failures, and stop only when the full sequence is green. | gronxb/ | 1.8k | — | ~572 | Automated safety check: Pass | Unknown | today |
| 39 | 39.Review Reviews a GitHub pull request for correctness, architecture, security, backward compatibility, and test coverage. | webern/ | 385 | — | ~2k | Automated safety check: Notes | Apache-2.0 | 2 days ago |
| 40 | A skill your agent uses when writing or modifying UE automated tests (Automation, CQTest, Functional, Gauntlet, LowLevel) with Rider MCP available. | JasonMa0012/ | 750 | — | ~2.1k | Automated safety check: Notes | Unknown | 23 days ago |
| 41 | Diagnoses and fixes failing GitHub Actions runs by identifying the failure, fetching only the relevant logs, finding the root cause and reproducing it locally. | ruby-git/ | 1.8k | — | ~1.9k | Automated safety check: Pass | MIT | 9 days ago |
| 42 | Embed a local image file into an existing GitHub PR — either in the PR body or as a comment. | bikeindex/ | 308 | — | ~1.9k | Automated safety check: Pass | AGPL-3.0 | today |
| 43 | A skill your agent uses when tests have race conditions, timing dependencies, or inconsistent pass/fail behavior - replaces arbitrary timeouts with condition polling to wait for actual state… | sandgardenhq/ | 137 | 3 repos | ~933 | Automated safety check: Pass | Unknown | 20 days ago |
| 44 | Helps build, test and extend the Qualcomm AI Engine Direct (QNN) backend in ExecuTorch, with routes for new ops, model export, Buck-vs-CMake parity fixes and per-layer accuracy debugging. | pytorch/ | 5.1k | — | ~1.8k | Automated safety check: Pass | Unknown | yesterday |
| 45 | 45.Diagnose Investigate unexpected behavior and mysterious bugs. An agent skill from avibebuilder/claude-prime. | avibebuilder/ | 120 | — | ~1.2k | Automated safety check: Pass | MIT | 4 mo ago |
| 46 | 46.Skill GitHub Handles GitHub issues and pull requests of the project with the gh CLI - list, triage, analyze an issue, review a PR, diagnose a CI failure, comment, label, close or merge. | Dolibarr/ | 7.7k | — | ~1.6k | Automated safety check: Pass | MIT | today |
| 47 | 47.Release App Cut and publish a new NarraCat-app version — bump the version number, build the Windows x64 package in CI, package + sign + notarize the macOS build locally, and publish both platforms into one… | yannikzz/ | 105 | — | ~1.6k | Automated safety check: Notes | AGPL-3.0 | 4 days ago |
| 48 | Classifies a failing test, typecheck or CI job before any code changes, by recording the failure and running a clean control to show whether it was already broken. | different-ai/ | 24k | — | ~779 | Automated safety check: Pass | Unknown | today |