Topic · DevOps & Cloud
Best runbooks and postmortems skills for Claude Code, Codex and other agents.
- skills
- 277
- official
- 11
Runbooks and postmortems skills, ranked
Ranked by score. Sort bymost stars,trending,newest,recently updated
| # | Skill | Repository | Stars | Used in | Tokens | Auto-check | Licence | Updated |
|---|---|---|---|---|---|---|---|---|
| 1 | Write or rewrite technical text with the rules of ASD-STE100 Simplified Technical English so it is clear, unambiguous, and free of AI slop. | moeru-ai/ | 50k | 2 repos | ~4.6k | Automated safety check: Pass | MIT | today |
| 2 | Runbook for assessing and executing a Mole CLI release: distribution channels, pre-flight checks, capital-V tags, build artifacts and the handoff to curated release notes. | tw93/ | 69k | — | ~2.5k | Automated safety check: Pass | GPL-3.0 | today |
| 3 | Track investment theses across their lifecycle — from screening idea to closed position with postmortem. | tradermonty/ | 3k | 2 repos | ~4.3k | Automated safety check: Pass | MIT | yesterday |
| 4 | Author or scope a first-party Nx migration. An agent skill from nrwl/nx. | nrwl/ | 29k | — | ~12k | Automated safety check: Notes | MIT | today |
| 5 | 5.Sepia Make AI-generated writing read as human-written, in fiction and in professional prose. | Nanako0129/ | 3k | — | ~3.6k | Automated safety check: Pass | MIT | yesterday |
| 6 | A skill your agent uses when a change is non-trivial by DSH standards (behavior, architecture, cross-file contracts, process/tooling, testing strategy, or on-disk/wire/config formats), when choosing… | czm15053/ | 477 | — | ~1.9k | Automated safety check: Pass | No licence | 15 days ago |
| 7 | Minimal starter runbook for cloud agents to install dependencies, run packages, execute tests, and troubleshoot the Scalar monorepo quickly. | scalar/ | 16k | — | ~1.4k | Automated safety check: Pass | MIT | today |
| 8 | Walks an agent through upgrading the OpenRig CLI and daemon one observed step at a time, keeping live seats alive and reconciling managed plugin files. | mvschwarz/ | 5.5k | — | ~2.9k | Automated safety check: Pass | Apache-2.0 | today |
| 9 | Research, create, improve, migrate, evaluate, package, install-check, govern, and safely publish qiaomu-flavored agent skills from workflows, prompts, transcripts, docs, SOPs, runbooks, scripts, or… | joeseesun/ | 383 | — | ~2.8k | Automated safety check: Pass | MIT | 2 mo ago |
| 10 | Runbook for publishing a GreptimeDB version: pick the release branch, verify the Cargo version, then tag, create the GitHub release and open the docs note PR. | GreptimeTeam/ | 6.7k | — | ~1.4k | Automated safety check: Pass | Apache-2.0 | 2 days ago |
| 11 | Write or rewrite text in plain, layman-readable English in the spirit of ASD-STE100 Simplified Technical English: short sentences, active voice, simple tenses, one word one meaning, condition before… | ropensci/ | 104 | 3 repos | ~2k | Automated safety check: Pass | MIT | 15 days ago |
| 12 | 12.Statem A skill your agent uses when a long coding or research task should be managed with statem state-machine runbooks, including creating specs, starting or resuming runs, checking current state… | henryqin1997/ | 1.3k | — | ~1.2k | Automated safety check: Pass | Apache-2.0 | 1 mo ago |
| 13 | 13.Make Trace Turn any source that describes how a kind of task gets done (a SKILL.md, a chat log, a runbook, plain prose) into a runnable Morph trace. | AIScientists-Dev/ | 485 | — | ~3.7k | Automated safety check: Pass | MIT | 4 mo ago |
| 14 | 14.Post Mortem Write the canonical engineering record of a fixed bug — root cause, mechanism, fix, validation, and how it slipped through. | thananon/ | 3.2k | — | ~3.4k | Automated safety check: Pass | No licence | 3 mo ago |
| 15 | Run a structured after-action review (postmortem, retrospective) on a launch, incident, or completed project to capture timeline, root cause analysis, contributing factors, and actionable lessons. | rampstackco/ | 935 | 1 repo | ~2.5k | Automated safety check: Pass | MIT | today |
| 16 | 16.CLI Release Runbook for releasing the executor CLI package (stable and beta). | UsefulSoftwareCo/ | 4.1k | — | ~2k | Automated safety check: Pass | MIT | today |
| 17 | Source-of-truth runbook for preparing this Vinext Cloudflare Workers SaaS template for production deployment. | LubomirGeorgiev/ | 786 | — | ~5.9k | Automated safety check: Notes | MIT | today |
| 18 | 18.Li Audit Post-mortem on what the user has already published - which posts actually worked, why, and what to stop doing. | Jakeschincariol/ | 1.4k | — | ~854 | Automated safety check: Pass | MIT | 20 days ago |
| 19 | 19.Audit Flow Interactive system flow tracing across CODE, API, AUTH, DATA, NETWORK layers with SQLite persistence and Mermaid export. | zebbern/ | 4.6k | — | ~4.2k | Automated safety check: Pass | MIT | today |
| 20 | Write a blameless incident postmortem under postmortems/ following the Google SRE shape — evidence-based timeline, trigger vs root cause vs symptom, contributing factors, what went well, and… | inkeep/ | 4.4k | — | ~4.4k | Automated safety check: Pass | GPL-3.0 | today |
| 21 | 21.Code Review Run CodeRabbit CLI reviews, retrieve saved local or GitHub PR fix prompts, and interpret CodeRabbit authentication and review output. | coderabbitai/ | 188 | — | ~1.9k | Automated safety check: Pass | MIT | today |
| 22 | Deploy Sablier protocols to a new EVM chain. An agent skill from sablier-labs/evm-monorepo. | sablier-labs/ | 353 | — | ~3.8k | Automated safety check: Notes | Unknown | 13 days ago |
| 23 | 23.Use Statem Manage long Claude Code work with statem state-machine runbooks. | henryqin1997/ | 1.3k | — | ~1.1k | Automated safety check: Pass | Apache-2.0 | 1 mo ago |
| 24 | 24.Okf Author, maintain, and consume Open Knowledge Format (OKF) knowledge bundles — portable markdown + YAML frontmatter that both humans and agents read. | scaccogatto/ | 407 | — | ~2.1k | Automated safety check: Notes | MIT | 8 days ago |
| 25 | Write well-formatted notes to the atmosphere-vault Obsidian knowledge base. | Atmosphere/ | 3.8k | — | ~2.6k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 26 | 26.Md2html Convert long-form Markdown (plan, spec, system design, RFC, runbook, postmortem, brainstorm, notes) into a single self-contained HTML page with Mermaid diagrams, step timelines, callouts, sidebar TOC. | haidang1810/ | 421 | — | ~3k | Automated safety check: Pass | MIT | 4 mo ago |
| 27 | 27.Fable Soul A skill your agent uses when a session or weaker model needs the Fable judgment layer loaded; when installing, refreshing, or syncing the soul across machines, runtimes, or global instruction files… | akseolabs-seo/ | 107 | — | ~1.1k | Automated safety check: Pass | MIT | 3 mo ago |
| 28 | Writes, repairs and copyedits project docs against the Google developer documentation style guide, verifying claims in the repository before stating them. | bgauryy/ | 946 | — | ~2k | Automated safety check: Pass | MIT | 3 days ago |
| 29 | 29.Oncall Pigweed oncall rotation runbooks and maintenance workflows (such as rolling CIPD client tools for b/315378787). | pigweed-project/ | 547 | — | ~992 | Automated safety check: Pass | Apache-2.0 | today |
| 30 | 30.DB Ops Sop Database operations runbook — backup, recovery, performance tuning, troubleshooting. | OpenDCAI/ | 406 | — | ~388 | Automated safety check: Pass | Apache-2.0 | 17 days ago |
| 31 | Loads one entity-linked Skill through UModel, follows its inline SKILL.md and applies attached knowledge items under their apply policies. | alibaba/ | 412 | — | ~1k | Automated safety check: Pass | Unknown | 13 days ago |
| 32 | Maintain bounded, durable AI project memory in a repository's docs/ai/ pack. | tudoumashu/ | 411 | — | ~1.1k | Automated safety check: Pass | MIT | 29 days ago |
| 33 | Think like a product manager before changing React Doctor's public surface — CLI commands/flags, the 0–100 score, config (doctor.config.), the JSON report schema, package APIs… | millionco/ | 15k | — | ~3.1k | Automated safety check: Pass | Unknown | today |
| 34 | 34.Ig Audit Post-mortem on what the user has already posted - which reels actually worked, why, and what to stop making. | Jakeschincariol/ | 458 | — | ~1.1k | Automated safety check: Pass | MIT | 24 days ago |
| 35 | Produce a self-contained HTML artifact instead of a markdown document when the content benefits from spatial layout, color, real diagrams, interactivity, or a round-trip editor. | dogum/ | 150 | — | ~2.1k | Automated safety check: Pass | Apache-2.0 | 22 days ago |
| 36 | Create and maintain consistent hand-drawn architecture visuals for pi-runbook. | Pr1p/ | 143 | — | ~1.4k | Automated safety check: Pass | No licence | 1 mo ago |
| 37 | Review closed trades, partial exits, and monthly trade aggregates for process adherence, risk discipline, execution quality, and evidence-based trading behavior patterns. | tradermonty/ | 3k | 1 repo | ~2.4k | Automated safety check: Pass | MIT | yesterday |
| 38 | Before answering any technical question, code request, architecture decision, or factual claim, call searchknowledge to check the local corpus. | lyonzin/ | 290 | — | ~1.4k | Automated safety check: Pass | MIT | 3 days ago |
| 39 | Design, validate, and govern fail-closed customer-activation automations that use an Outbox/worker pattern. | Ali-Marandi/ | 107 | — | ~1.9k | Automated safety check: Pass | MIT | 1 mo ago |
| 40 | Documentation templates for ADRs, runbooks, architecture docs, and knowledge transfer documents. Use when creating Architecture Decision Records… | bybren-llc/ | 421 | — | ~1.2k | Automated safety check: Pass | MIT | 2 mo ago |
| 41 | Plan, smoke-test, execute, checkpoint, publish, audit, and reproduce full ShellBench native benchmark campaigns across OpenClaw, Hermes, Codex, and Claude Code, including model and reasoning… | openclaw/ | 141 | — | ~1.1k | Automated safety check: Pass | MIT | today |
| 42 | Guideline-authoring craft for instruction SoTs (.claude/ guides, CLAUDE.md, plan files) — operative rules vs recital, MUST/SHOULD/MAY force tiers, pruning, bloat control, blind review protocol… | alfadur7/ | 170 | — | ~2.4k | Automated safety check: Pass | MIT | 2 days ago |
| 43 | Full runbook for cutting a Blanc desktop release — scripts/release.sh mechanics and its required BLANCRELEASE env vars, macOS notarization via 1Password, the Touch ID provisioning profile and… | bnfy/ | 105 | — | ~2.9k | Automated safety check: Notes | MIT | today |
| 44 | Release/deploy workflow for Agent Sessions (Sparkle appcast + GitHub release). | jazzyalex/ | 892 | — | ~817 | Automated safety check: Pass | MIT | today |
| 45 | Post-mortem analysis of CI failures across recent PRs in dotnet/macios. | dotnet/ | 2.9k | — | ~7.8k | Automated safety check: Pass | Unknown | today |
| 46 | 46.Code Review Review pull requests, diffs, and code changes in z-shell repositories against the bundled organization criteria and the repository's own contracts and checks, verifying each finding before reporting… | z-shell/ | 121 | — | ~1.6k | Automated safety check: Pass | MIT | today |
| 47 | A skill your agent uses when the user asks for overloaded-mode or burnout-mode, describes burnout or being burned out, or is overwhelmed, frozen, overcommitted, burnout-adjacent, or unable to decide… | softcane/ | 116 | — | ~1.4k | Automated safety check: Pass | MIT | 2 mo ago |
| 48 | Writes PR descriptions, changelog entries, release notes and postmortems from the actual commits and diff, and saves each one in the right place. | jsmastery-pro/ | 1.4k | — | ~2.4k | Automated safety check: Notes | MIT | 1 mo ago |
Questions, answered from the data.
What is the best runbooks and postmortems skill?
Simple English from moeru-ai/airi ranks first of the 277 runbooks and postmortems skills listed here, with the highest score: its repository has 50k GitHub stars, 2 other GitHub owners carry a copy, its SKILL.md loads about 4.6k tokens and it passes the automated safety check with no findings. Next come Mole CLI Release Flow and Trader Memory Core.
Which runbooks and postmortems skills are official?
11 of the 277 runbooks and postmortems skills are official, published by the vendor's own GitHub organization: Macios CI Postmortem, Sdaf Sovereign Cloud, Campaign Postmortem, Incident Postmortem, Tao Validate Recipe Transfer and 6 more.
How are these skills ranked?
By Skill Navigator score, which combines the GitHub stars of the skill's repository (shared across that repo's skills and discounted for large collections), how many other GitHub owners carry a copy of the skill, and automated SKILL.md quality checks, minus penalties for safety-check warnings and for each further skill from the same repository. Skills that fail the safety check are not listed.
Explore related skills
Category
More topics in DevOps & Cloud
- Deployment1,152
- CI/CD921
- Containers723
- Observability562
- Container orchestration519
- Infrastructure as code351
- Monitoring and alerting333
- Secrets management318
- Incident response271
- Cloud networking230
- Backup and disaster recovery170
- Site reliability engineering150
- Cloud architecture114
- MLOps101
- Cloud cost optimization93
- GitOps87
- Linux administration75
- Platform engineering51
- Chaos engineering28