Topic · Development

Best root cause analysis skills for Claude Code, Codex and other agents.

Skills that trace a failure back to its underlying cause rather than patching symptoms.
skills
610
official
56

Root cause analysis skills, ranked

Ranked by score. Sort bymost stars,trending,newest,recently updated

Root cause analysis skills, ranked
#SkillRepositoryStarsUsed inTokensAuto-checkLicenceUpdated
1

Investigates a session where Superpowers went wrong, reads the transcripts on disk and produces an evidence-cited report, optionally prepared as a bug report for the maintainers.

obra/superpowers296k3 repos~1.7kAutomated safety check: PassMITyesterday
2

Digs into why code is shaped the way it is by checking git history, pull requests and connected tools in parallel, then reporting a cited read on the tradeoffs.

cursor/plugins10k9 repos~2.6kAutomated safety check: PassNo licenceyesterday
3

Investigates failing Pester tests in PowerShell CI jobs by following a six-step workflow from pull request status to documented fix recommendations.

PowerShell/PowerShell56k—~5.1kAutomated safety check: PassMITyesterday
4

Decides whether an OpenLogi device problem on macOS is a privacy-permission (TCC) problem, using agent log lines, and says which identity needs which grant.

AprilNEA/OpenLogi23k—~2.5kAutomated safety check: NotesApache-2.04 days ago
5

Forms a two-agent committee with contrasting profiles to analyze a stuck problem in parallel, reconcile their views and return a consensus plan without editing files.

getpaseo/paseo20k1 repo~496Automated safety check: PassUnknownyesterday
6

Mines local Copilot CLI session logs for dotnet/maui to rank costly or failing runs, tag recurring failure modes, propose repo edits and emit guard evals.

dotnet/maui23k—~3.4kAutomated safety check: PassMITyesterday
7

Investigates a stubbornly failing Playwright test as a possible product bug, using error output, screenshots, traces and server code, and writes a structured bug report.

appsmithorg/appsmith41k—~1.5kAutomated safety check: PassApache-2.0yesterday
8

Investigates suspected bugs in the daisyUI monorepo through read-only analysis, then writes a decision-ready fix plan in tmp/bugs without changing any product code.

saadeghi/daisyui43k—~2.3kAutomated safety check: PassMIT8 days ago
9

Investigates bugs, errors and stack traces in phases and requires a root-cause hypothesis to be confirmed before any fix is written.

garrytan/gstack136k—~1.4kAutomated safety check: PassMITyesterday
10

Fixes failing Playwright specs by reading the error, classifying the cause in the test code and applying corrections that follow project conventions.

appsmithorg/appsmith41k—~1.3kAutomated safety check: PassApache-2.0yesterday
11

Traces a bug through a code knowledge graph, following callers, callees and execution flow before opening source files, within a small token budget.

tirth8205/code-review-graph32k1 repo~287Automated safety check: PassMITyesterday
12

Answers questions about a past agent run from its recording, using causal graphs and replay, instead of reconstructing events from memory.

iflytek/skillhub5.2k4 repos~3kAutomated safety check: PassApache-2.06 days ago
13

Diagnoses failing CI pipelines and tests, deciding first whether the test or the implementation is at fault, and hands hard cases to a dedicated fixer subagent.

Chachamaru127/claude-code-harness3.2k1 repo~1.1kAutomated safety check: NotesMIT3 days ago
14

Triages findings from a Strix pentest by severity, fixes each root cause with a minimal change, and re-runs Strix to confirm the exploit no longer works.

usestrix/strix67k—~1.5kAutomated safety check: PassApache-2.0yesterday
15

Investigates TiDB plan or test-result diffs that the change does not explain, ruling out failpoint setup and merge effects before expected outputs are updated.

pingcap/tidb41k—~498Automated safety check: PassApache-2.0yesterday
16

Diagnoses surprising LoopX behavior, such as stale recommendations or tiny progress, assigns it to the responsible layer and repairs it at the lowest durable level.

loopx-project/loopx6.2k—~2.2kAutomated safety check: PassApache-2.0yesterday
17

Guided reverse engineering workflow for binaries, firmware, mobile apps, scripts, document samples, protocol captures, and unknown artifacts.

lingbol088-spec/reverse-flow-skill935—~2.4kAutomated safety check: PassMIT2 mo ago
18

Fix or implement a tracker issue end to end from a single command — takes an issue id or a plain problem description (filed first via om-prepare-issue), classifies, then drives the bug autofix chain…

go-musicfox/go-musicfox2.6k1 repo~5kAutomated safety check: NotesGPL-3.01 mo ago
19

Diagnoses where an agent failed across runs and turns the findings into new skills, system prompt patches and knowledge entries, using the A-Evolve loop.

aiming-lab/AutoResearchClaw15k—~1.8kAutomated safety check: PassMIT1 mo ago
20

Applies a four-phase debugging routine that finds the root cause of a bug or failing test before any fix is written.

ChrisWiles/claude-code-showcase6.1k3 repos~1.2kAutomated safety check: PassNo licence9 mo ago
21

Applies a stop-the-line rule and a step-by-step triage when tests fail, builds break or something stops working, aiming at the root cause instead of guesses.

addyosmani/agent-skills102k1 repo~2.6kAutomated safety check: PassMIT4 days ago
22

Pushes an agent to keep verifying and changing approach after repeated failures, using a diagnosis line, evidence-based completion and confirmation before risky edits.

tanweai/pua20k—~502Automated safety check: PassMIT28 days ago
23

Investigates past Kubernetes incidents from Kubeshark traffic snapshots: takes captures, dissects API calls, extracts PCAPs and compares traffic over time.

kubeshark/kubeshark12k—~5.3kAutomated safety check: PassApache-2.07 days ago
24

Diagnoses a failed GreptimeDB fuzz CI job by pulling its GitHub Actions logs and fuzz artifacts, then matching the evidence to the local source code.

GreptimeTeam/greptimedb6.7k—~4.4kAutomated safety check: PassApache-2.03 days ago
25

Investigates a failed Opik end-to-end test from CI, TestOps or a local run, decides regression versus flake, and proposes a fix without editing tests.

comet-ml/opik22k—~1.8kAutomated safety check: PassApache-2.0yesterday
26

Master systematic debugging techniques, profiling tools, and root cause analysis to efficiently track down bugs across any codebase or technology stack.

sangrokjung/claude-forge84912 repos~3.1kAutomated safety check: PassMIT1 mo ago
27

Debugging guide for Extempore covering its three layers, compilation paths, startup sequence and the batch, eval and interactive modes used to isolate JIT problems.

digego/extempore1.5k—~4.4kAutomated safety check: PassNo licence13 days ago
28

Conduct evidence-backed Happier code, plan-completeness, session, worktree, feature, commit, branch, PR, codebase, and release-readiness reviews with affected-corridor analysis, high-confidence…

happier-dev/happier1.9k—~4.5kAutomated safety check: PassMITtoday
29

Condensed debugging method and tool recipes for Unix, Python and PyTorch programs: crashes, hangs, segfaults, wrong output, CUDA OOM, NaN values and slowness.

stas00/the-art-of-debugging1.7k—~6.1kAutomated safety check: NotesCC-BY-SA-4.0yesterday
30

A harness that makes Opus (or any Claude model) behave like Fable — it enforces seeing a task through to the end, with evidence and verification, as procedure.

fivetaku/fablize895—~1.6kAutomated safety check: PassMIT3 mo ago
31

Systematic evidence-based debugging using runtime logs. An agent skill from millionco/expect.

millionco/expect3.6k—~2.6kAutomated safety check: PassUnknown5 mo ago
32

Fixes an OpenROAD bug from a GitHub issue or error code: finds the root cause, implements the fix, adds a regression test and prepares a signed-off commit.

The-OpenROAD-Project/OpenROAD3.2k—~784Automated safety check: PassBSD-3-Clauseyesterday
33

A skill your agent uses when asked to run Design Error Detection (quick defect scan), find design errors in a Simulink model, perform root cause analysis on DED findings, fix division-by-zero…

matlab/simulink-agentic-toolkit1.2k—~2.1kAutomated safety check: PassUnknown7 days ago
34

Write the canonical engineering record of a fixed bug — root cause, mechanism, fix, validation, and how it slipped through.

thananon/9arm-skills3.2k—~3.4kAutomated safety check: PassNo licence3 mo ago
35

Walks through open Sentry issues for the tooll3 project, latest first, proposing a fix for each and committing them one at a time with your review between.

tixl3d/tixl5.1k—~2.1kAutomated safety check: NotesMITyesterday
36

A skill your agent uses when encountering any bug, test failure, or unexpected behavior, before proposing fixes - four-phase framework (root cause investigation, pattern analysis, hypothesis…

ed3dai/ed3d-plugins2503 repos~2.4kAutomated safety check: PassNo licence1 mo ago
37

Guides systematic root-cause debugging. An agent skill from abashev/vfs-s3.

abashev/vfs-s31066 repos~2.6kAutomated safety check: PassApache-2.07 days ago
38

Run a structured after-action review (postmortem, retrospective) on a launch, incident, or completed project to capture timeline, root cause analysis, contributing factors, and actionable lessons.

rampstackco/claude-skills9351 repo~2.5kAutomated safety check: PassMITtoday
39

Researches code with evidence: traces callers, imports and cross-repo links, diagnoses failures and reports findings with exact file and line references and a confidence label.

bgauryy/octocode946—~1.5kAutomated safety check: PassMIT4 days ago
40

Triage failing GitHub PR checks: list failures with gh, fetch capped Actions logs, skip non-Actions checks, and summarize root cause.

Mentra-Community/MentraOS2.4k—~582Automated safety check: PassApache-2.0today
41

Analyze MSBuild binary logs to diagnose build failures. An agent skill from microsoft/testfx.

microsoft/testfx1k3 repos~730Automated safety check: PassMITtoday
42
42.Trx AnalysisOfficial

Parse and analyze Visual Studio TRX test result files. An agent skill from microsoft/vstest.

microsoft/vstest969—~1.8kAutomated safety check: PassMITyesterday
43

矛盾分析法:把复杂问题拆成若干对立面,找出规定其他矛盾的主要矛盾及其主要方面,判定对抗性 / 非对抗性,并据此选择处理方式。当问题头绪多、多个因素互相牵制、优先级不清、根因不明、反复修不好、trade-off 说不清时触发;直接执行类任务或用户已定方案时不触发。

HughYau/qiushi-skill3.8k—~485Automated safety check: PassMIT6 days ago
44

After a bug is found, traces its root cause and feeds a new testable invariant back into the project spec so the bug class can't recur.

JuliusBrussee/cavekit1.1k—~653Automated safety check: PassMIT1 mo ago
45

Forces a one-sentence, evidence-backed root cause before any fix is applied, and gates when a diagnosis session is even allowed to touch code.

tw93/Waza7.2k—~4.3kAutomated safety check: PassMITyesterday
46

A skill your agent uses for ANY bug, error, crash, wrong output, loss divergence, gradient explosion, test failure, CUDA error, distributed training hang, checkpoint load failure, or unexpected…

ByteDance-Seed/VeOmni2.2k—~2.8kAutomated safety check: PassApache-2.07 days ago
47

Investigates a service incident to its root cause by querying a UModel object graph alongside metrics, logs, topology and recent deployments.

alibaba/UnifiedModel412—~1.9kAutomated safety check: PassUnknown14 days ago
48

Investigate unexpected behavior and mysterious bugs. An agent skill from avibebuilder/claude-prime.

avibebuilder/claude-prime1201 repo~1.2kAutomated safety check: PassMIT4 mo ago

Questions, answered from the data.

What is the best root cause analysis skill?

Diagnosing Superpowers Sessions from obra/superpowers ranks first of the 610 root cause analysis skills listed here, with the highest score: its repository has 296k GitHub stars, 3 other GitHub owners carry a copy, its SKILL.md loads about 1.7k tokens and it passes the automated safety check with no findings. Next come Code Design Rationale Investigator and Pester Failure Analysis.

Which root cause analysis skills are official?

56 of the 610 root cause analysis skills are official, published by the vendor's own GitHub organization: Code Design Rationale Investigator, Copilot Session Failure Analysis, Binlog Failure Analysis, Trx Analysis, Triagebot Action Bug Triage and 51 more.

How are these skills ranked?

By Skill Navigator score, which combines the GitHub stars of the skill's repository (shared across that repo's skills and discounted for large collections), how many other GitHub owners carry a copy of the skill, and automated SKILL.md quality checks, minus penalties for safety-check warnings and for each further skill from the same repository. Skills that fail the safety check are not listed.