Search

Failing and flaky tests

553 skills found, page 10.
Search results
#SkillRepositoryStarsUsed inTokensAuto-checkLicenceUpdated
433

Content-assertion sweep after editing SKILL.md files. An agent skill from ZaxbyHub/opencode-swarm.

ZaxbyHub/opencode-swarm494—~532Automated safety check: PassMITyesterday
434

Weekly. An agent skill from markfulton/ai-employees.

markfulton/ai-employees543—~14kAutomated safety check: PassMITyesterday
435

Write or revise tests with an emphasis on behavior, regression coverage, pytest style, and avoiding "slop tests." Use when adding tests, fixing failing tests, reviewing test quality, or deciding…

open-thoughts/OpenThoughts-Agent301—~498Automated safety check: PassApache-2.013 days ago
436
436.Ut CheckOfficial

Analyze UT (unit test) results for a torch-xpu-ops PR. An agent skill from intel/torch-xpu-ops.

intel/torch-xpu-ops115—~1.6kAutomated safety check: PassApache-2.0yesterday
437

A skill your agent uses when an agent says a test, VM, proof, evaluator, or release gate is overloaded, too strict, blocking staging, or causing false positives; split checks by profile, measure the…

AnastasiyaW/codex-claude-code-config154—~1kAutomated safety check: PassMIT2 days ago
438

Debug and fix a bug by starting from one concrete failure, forming a falsifiable local hypothesis, making the smallest grounded edit, and immediately running focused validation.

lge-ros2/cloisim176—~1.2kAutomated safety check: PassMIT5 days ago
439

Inspect GitHub PR for CI failures, merge conflicts, update-branch requirements, reviewer comments, change requests, and unresolved review threads.

akiojin/unity-cli108—~4.6kAutomated safety check: PassApache-2.02 days ago
440

Analyzes DevOps Center test failures and Code Analyzer violations in plain language — failure category, offending file/class/method/line, rule violated, fix direction, and prioritized improvement…

forcedotcom/sf-skills1.1k—~1.4kAutomated safety check: PassApache-2.0yesterday
441

Orchestrates test suite execution with parallel sharding, intelligent retry, and real-time reporting across Jest, Vitest, and Playwright.

proffesor-for-testing/agentic-qe495—~1.2kAutomated safety check: PassMIT2 days ago
442

Design and implement effective test automation with proper pyramid, patterns, and CI/CD integration.

proffesor-for-testing/agentic-qe495—~1.4kAutomated safety check: PassMIT2 days ago
443

Guide for disabling Chromium tests (C++ or WebUI). An agent skill from nwjs/chromium.src.

nwjs/chromium.src160—~790Automated safety check: PassBSD-3-Clauseyesterday
444

Generate and run tests for Adobe App Builder actions and UI components.

adobe/skills197—~2.8kAutomated safety check: PassApache-2.0yesterday
445

Identifies non-deterministic or unreliable tests through static code analysis and test result analysis.

ArabelaTso/Skills-4-SE253—~2kAutomated safety check: PassApache-2.01 mo ago
446

Automatically updates a codebase to a new language version, framework version, or library update while ensuring all tests still pass.

ArabelaTso/Skills-4-SE253—~1.8kAutomated safety check: PassApache-2.01 mo ago
447

Write and evaluate effective Python tests using pytest. An agent skill from benchflow-ai/skillsbench.

benchflow-ai/skillsbench1.8k—~1.3kAutomated safety check: PassApache-2.02 mo ago
448

Test-first development route for TDD, writing failing tests first, RED - GREEN - REFACTOR, and behavior-changing feature/bug/refactor work.

foryourhealth111-pixel/Vibe-Skills3.6k—~455Automated safety check: PassApache-2.01 mo ago
449

Comprehensive testing guide for Cloudflare Workers using Vitest and @cloudflare/vitest-pool-workers.

secondsky/claude-skills227—~5.3kAutomated safety check: PassMIT13 days ago
450

Testing strategy: pyramid, AAA, mocks/fakes/stubs, flaky tests, coverage.

softspark/ai-toolkit179—~1.6kAutomated safety check: PassApache-2.03 days ago
451
451.Trace

A skill your agent uses when encountering bugs, test failures, runtime errors, unexpected behavior, broken builds, or "this doesn't work" reports.

ccplugins/awesome-claude-code-plugins970—~1.4kAutomated safety check: PassApache-2.02 mo ago
452

Write or repair Elixir tests with ExUnit, sandbox isolation, async reliability, Mox, ExMachina, and LiveViewTest.

oliver-kriska/claude-elixir-phoenix565—~750Automated safety check: PassMIT6 days ago
453

Minimum-requirements checklist for any change — code or docs-only.

oocx/tfplan2md174—~1.8kAutomated safety check: PassMITyesterday
454

Red-green-refactor scaffold for building new features with TDD.

gustavscirulis/snapgrid1161 repo~2.4kAutomated safety check: NotesUnknown5 mo ago
455

Isolate complex Adobe failures across DNS/TLS, auth, entitlement, schema, async state, storage, Runtime, and downstream layers without leaking credentials or content.

jeremylongshore/tons-of-skills-marketplace2.8k—~1.1kAutomated safety check: PassMITyesterday
456

Isolate difficult Canva Connect failures without rotating credentials, replaying writes, or leaking customer data.

jeremylongshore/tons-of-skills-marketplace2.8k—~1kAutomated safety check: PassMITyesterday
457

Deep debugging for Figma API issues: network analysis, response inspection, and support escalation.

jeremylongshore/tons-of-skills-marketplace2.8k—~1.9kAutomated safety check: PassMITyesterday
458

Debug complex Shopify API issues using cost analysis, request tracing, webhook delivery inspection, and GraphQL introspection.

jeremylongshore/tons-of-skills-marketplace2.8k—~1.3kAutomated safety check: PassMITyesterday
459
459.Trace

A skill your agent uses when encountering bugs, test failures, runtime errors, broken builds, or "this doesn't work" reports.

jeremylongshore/tons-of-skills-marketplace2.8k—~5kAutomated safety check: PassMITyesterday
460

Advanced debugging for hard-to-diagnose Vercel issues including cold starts, edge errors, and function tracing.

jeremylongshore/tons-of-skills-marketplace2.8k—~2.1kAutomated safety check: PassMITyesterday
461

End-to-end testing with Playwright: test generation, page objects, locator strategy, flaky- test diagnosis, visual regression, and CI integration.

borghei/Claude-Skills891—~2kAutomated safety check: PassMIT4 days ago
462

Diagnose a failing CI workflow run (lint, markdown lint, build, or unit tests) — identify which job failed, the cause, and a concrete fix

adamayoung/TMDb178—~1.6kAutomated safety check: PassApache-2.07 days ago
463

Monitor a PR and fix CI or review feedback. An agent skill from BuilderIO/agent-native.

BuilderIO/agent-native7.1k—~8.5kAutomated safety check: PassNo licenceyesterday
464

Diagnose a PR, branch, run, or GitHub Actions URL from CI evidence and produce a fix plan without changing code.

Terry-Mao/AICodingFlow167—~447Automated safety check: PassMIT8 days ago
465
465.Openai Gh Fix CIOfficial

A skill your agent uses when a user asks to debug or fix failing GitHub PR checks that run in GitHub Actions; use gh to inspect checks and logs, summarize failure context, draft a fix plan, and…

trailofbits/skills-curated513—~955Automated safety check: NotesCC-BY-SA-4.02 mo ago
466

Iterate on a GitHub pull request — drive it through CI, code review, and QA until it is merge-ready.

OpenHands/extensions163—~3.8kAutomated safety check: PassMIT2 days ago
467

Running and debugging tests in the Golem workspace. An agent skill from golemcloud/golem.

golemcloud/golem1.5k—~2kAutomated safety check: PassUnknownyesterday
468

Codex adapter for monitoring and fixing CI failures on opencode-swarm PRs.

ZaxbyHub/opencode-swarm494—~551Automated safety check: PassMITyesterday
469

A skill your agent uses when asked to trace, investigate, root-cause, reproduce, plan, fix, resolve, close, or prepare a PR for an issue, bug report, defect, regression, failing test, or confusing…

ZaxbyHub/opencode-swarm494—~383Automated safety check: PassMITyesterday
470

A skill your agent uses when asked to trace, investigate, root-cause, reproduce, plan, fix, resolve, close, or prepare a PR for an issue, bug report, defect, regression, failing test, or confusing…

ZaxbyHub/opencode-swarm494—~397Automated safety check: PassMITyesterday
471

run failing tests immediately after testgate failure. An agent skill from ZaxbyHub/opencode-swarm.

ZaxbyHub/opencode-swarm494—~332Automated safety check: PassMITyesterday
472
472.Fix

Fix bugs and broken behavior when there is enough evidence to act on a repair path.

avibebuilder/claude-prime120—~1kAutomated safety check: PassMIT4 mo ago
473

Deliver changes in this monorepo through Git-Spice, focused checks, Woodpecker CI, automated review, and PR evidence.

shepherdjerred/monorepo112—~962Automated safety check: PassGPL-3.0yesterday
474
474.Fix ReproduceOfficial

A skill your agent uses when asked to reproduce a bug, verify a nightly CI failure, or confirm a failure still exists on latest source.

intel/torch-xpu-ops115—~7.7kAutomated safety check: PassApache-2.0yesterday
475

A skill your agent uses whenever the design will encounter conditions that can't be fully predicted at design time — load spikes, edge-case content, real-world variability, hardware variation…

hashgraph-online/awesome-codex-plugins1.3k—~3.5kAutomated safety check: PassApache-2.0yesterday
476

Load when investigating a failing PR CI pipeline or checking PR health.

datadog-labs/agent-skills177—~2.7kAutomated safety check: PassMIT2 days ago
477

Systematically fix all failing tests after business logic changes or refactoring

NeoLabHQ/context-engineering-kit1.8k—~914Automated safety check: PassGPL-3.01 mo ago
478

Production incidents dashboard. An agent skill from davepoon/buildwithclaude.

davepoon/buildwithclaude3.6k—~1.8kAutomated safety check: NotesMIT2 days ago
479

Autonomous PR merge pipeline. An agent skill from davepoon/buildwithclaude.

davepoon/buildwithclaude3.6k—~3.1kAutomated safety check: NotesMIT2 days ago
480

Operate Playwright for browser automation end to end: author and debug E2E test suites (robust locators, network interception and mocking, parallel workers, accessibility snapshot checks), wire them…

magnus919/agent-skills115—~3.4kAutomated safety check: NotesMITyesterday