Topic · Agent Workflows

Best verification before completion skills for Claude Code, Codex and other agents.

Skills that make agents prove their work with checks before saying they are done.
skills
169
official
8

Verification before completion skills, ranked

Ranked by score. Sort bymost stars,trending,newest,recently updated

Verification before completion skills, ranked
#SkillRepositoryStarsUsed inTokensAuto-checkLicenceUpdated
1

Keeps a TSV decision log for long or unattended agent runs, one row per decision with what, why, evidence and result, so a reviewer can check the work later.

cursor/plugins10k9 repos~1.6kAutomated safety check: PassNo licenceyesterday
2

Verifies a delivery end to end by driving the real product on a CLI, web, desktop or iOS Simulator surface, capturing evidence and publishing a round with the lh CLI.

lobehub/lobehub83k—~9.7kAutomated safety check: PassApache-2.0today
3

A skill your agent uses when about to claim work is complete, fixed, or passing, before committing or creating PRs - requires running verification commands and confirming output before making any…

farm-fe/farm5.6k45 repos~1kAutomated safety check: PassMIT15 days ago
4

Keeps an agent focused on the requested work by applying a five-step ladder that checks for direct solutions, real gaps and speculative defenses before adding anything.

lennney/stop-that-shit2.5k1 repo~2kAutomated safety check: PassMIT2 days ago
5

Adds short, pointed workplace reminders drawn from two Chinese essays on corporate culture, nudging the agent to prove results with evidence instead of polished reports.

tanweai/pua20k—~556Automated safety check: PassMIT28 days ago
6

Proves a Codewhale change in the real product: a stamped release build, an atomic local install, fresh-shell verification and manual QA that automated gates cannot cover.

codewhale-hq/Codewhale41k—~1.3kAutomated safety check: PassMITtoday
7

Explains why editing CUTLASS fused-MHA headers in ONNX Runtime can leave stale CUDA kernels after an incremental build, and how to force and verify a real rebuild.

microsoft/onnxruntime22k—~1.3kAutomated safety check: PassMITtoday
8

Records an annotated screen recording of the agent testing an app hands-on, then posts the video and a results summary to the PR and tracker issue.

michaelshimeles/skills1.3k1 repo~3.9kAutomated safety check: PassNo licence3 days ago
9

Read and write email through the Atomic Mail from an AI agent.

Atomic-Mail/atomic-mail-agentic2661 repo~2kAutomated safety check: PassMIT10 days ago
10

Pushes an agent to keep verifying and changing approach after repeated failures, using a diagnosis line, evidence-based completion and confirmation before risky edits.

tanweai/pua20k—~502Automated safety check: PassMIT28 days ago
11

The binding directive for ultrawork mode: evidence-first delivery, a light or heavy process tier, and constant use of memory during the work.

code-yeongyu/lazycodex3.7k—~6.9kAutomated safety check: PassMITyesterday
12

Delivers a change in thin vertical slices, each implemented, tested, verified and committed before the next, using vertical, contract-first or risk-first slicing.

addyosmani/agent-skills102k1 repo~2.3kAutomated safety check: PassMIT4 days ago
13

Runs an unattended iterate-until-verified loop in which a user-set verify command, not the agent's own claim, decides when the task is finished.

tanweai/pua20k—~1.1kAutomated safety check: PassMIT28 days ago
14

Executes a written implementation plan task by task in the current session, with a progress ledger, test-first gates and one fresh-context review at the end.

jnMetaCode/superpowers-zh8.3k—~2.5kAutomated safety check: PassMIT3 days ago
15

Runs the Antigravity CLI headlessly with the agy command to analyze a repository or carry out a coding task, then checks its JSON result and the diff.

CherryHQ/cherry-studio52k—~531Automated safety check: PassAGPL-3.0today
16

Confirms that a change to PlotJuggler 4 really works in the running app by proving the rebuild, launching with real data and measuring the result.

PlotJuggler/PlotJuggler6.2k—~956Automated safety check: PassMPL-2.06 days ago
17

SkillsBench task authoring — walk a contributor from idea to submission-ready task following CONTRIBUTING.md and the task-implementation rubric.

benchflow-ai/benchflow353—~4.5kAutomated safety check: PassApache-2.0yesterday
18

Researches code with evidence: traces callers, imports and cross-repo links, diagnoses failures and reports findings with exact file and line references and a confidence label.

bgauryy/octocode946—~1.5kAutomated safety check: PassMIT4 days ago
19

A skill your agent uses when an AI agent wants to participate in the Haidian Centennial Jing-Zhang AI Innovation Belt open call, follow changing materials and community discussion, generate or…

open-city-ai/haidian415—~9.4kAutomated safety check: PassNo licence1 mo ago
20

A skill your agent uses when about to claim a result, effect, significance, or that an analysis reproduces, before reporting or writing it up - requires running the analysis fresh and reading the…

K-Dense-AI/science-superpowers3481 repo~1.5kAutomated safety check: PassUnknown24 days ago
21

Decomposes a complex feature into tasks, dispatches parallel specialist agents with durable state, and supervises verification, QA review and retries.

first-fluke/oh-my-agent1.3k—~4.1kAutomated safety check: PassMITyesterday
22

A skill your agent uses when the user asks for an AWS architecture diagram — VPC/networking, event-driven, landing zone, multi-AZ, serverless pipeline, or any diagram built with AWS service icons.

sparklabx/drawio-ai-kit652—~1.6kAutomated safety check: PassMIT5 days ago
23

Checks generated or edited documentation against the source code, flagging invented symbols, outdated samples and unverifiable claims before publishing.

amElnagdy/guard-skills1.3k—~2.1kAutomated safety check: PassMIT3 mo ago
24
24.Advisor ModeOfficial

Adds a second, stronger model that the main agent consults before major decisions, when stuck and before finishing, controlled by /advisor commands.

cursor/plugins10k—~2.6kAutomated safety check: NotesNo licenceyesterday
25

A gate before merge: verify runs the real app against the spec, and review has a different model do a senior code review, without editing code.

jsmastery-pro/skills1.4k—~1.1kAutomated safety check: NotesMIT1 mo ago
26

Coordinates and recovers multi-stage Light research projects from a single passport file, with checkpoints, stale-work tracking and rerouting only when you approve.

Light0305/Light-skills640—~3.8kAutomated safety check: PassMIT3 mo ago
27

Runs an evidence-bound audit, daily or weekly review through the requirement-ledger CLI, binding scope and authority first and blocking the final report until a handoff check passes.

adand-91/gpt-6-astra-skill125—~1.5kAutomated safety check: PassMIT21 days ago
28

Turns a one-line idea into a task brief of up to 4000 characters that an autonomous agent can run through /goal, with measured baselines and cheat-resistant acceptance checks.

KKKKhazix/khazix-skills21k—~723Automated safety check: PassMIT6 days ago
29

Pushes an agent that keeps failing or gives up to exhaust every option, using harsh corporate-pressure wording, a diagnosis line and a proactivity checklist.

tanweai/pua20k—~3.3kAutomated safety check: PassMIT28 days ago
30

A skill your agent uses when the user asks for an Azure architecture diagram — VNet/networking, App Service, AKS, landing zone, multi-region, or any diagram built with Azure service icons.

sparklabx/drawio-ai-kit652—~1.6kAutomated safety check: PassMIT5 days ago
31

Applies a 12-stage verified workflow, from research to deploy, to non-trivial coding tasks, scaled to lightweight, standard or full mode by task size.

artemiimillier/bulletproof153—~3.5kAutomated safety check: PassMIT6 mo ago
32

Proves the Puppetmaster Jev transition behaves as specified: unset never uses the network, opt-in may skip the conflict auditor, and later versions only observe.

professorpalmer/Puppetmaster467—~525Automated safety check: PassMITtoday
33

Final checklist before handing work back in the herdr-auto-title repo: review the diff, run make check, apply the comment and AGENTS.md rules, then report.

kryptamine/herdr-auto-title237—~605Automated safety check: PassMITyesterday
34

Run the infrastructure health check and fix anything that fails

diet103/claude-code-infrastructure-showcase10k—~363Automated safety check: PassMIT2 mo ago
35

Read-only detector that compares SPEC.md with the code and reports invariant, interface and task drift grouped by severity, without changing anything.

JuliusBrussee/cavekit1.1k—~666Automated safety check: PassMIT1 mo ago
36

Scores an agent's finished work with a three-stage pipeline: free mechanical checks, an advisory semantic review, and an optional multi-model consensus vote.

Q00/ouroboros6.2k—~2.2kAutomated safety check: PassMITyesterday
37

Human handoff, Linear branch names, reviewable (non-draft) PRs, the cursor GitHub label, Claude, Greptile, and Codex review comments, preview test steps, proof of work posted on the GitHub PR, and…

langfuse/langfuse35k—~1.7kAutomated safety check: PassUnknowntoday
38

A work-discipline protocol that makes Opus 4.8 (or any non-frontier model) operate at Fable-5-grade quality.

cozytab/fable5-mode106—~4.8kAutomated safety check: PassMIT2 mo ago
39

A skill your agent uses when the user wants to set up / scaffold / install a file-based multi-agent orchestration system in a folder.

netwaif/multi-agent-starter106—~629Automated safety check: PassMIT1 mo ago
40

Applies Procoder's senior-developer discipline in a repository: run the commit gate, format through the binary and work through specs, plans and todos.

azrtydxb/procoder211—~3.7kAutomated safety check: PassApache-2.09 days ago
41

A required checklist for after code changes: validate each changed layer, check that docs and other layers stay in sync, and test before committing or opening a PR.

ZeroDeng01/sublinkPro1.7k—~4.4kAutomated safety check: PassMIT2 days ago
42

Runs a math modeling contest pipeline for CUMCM and MCM/ICM entries in Codex CLI, with git checkpoints, verified solver code and human review at each stage.

RealSeaberry/AutoMCM-Pro258—~1.6kAutomated safety check: PassMIT27 days ago
43

Pushes an agent that keeps failing, gives up or claims unverified success into a diagnosis, evidence and verification loop, with a Pi extension for persistent mode.

tanweai/pua20k—~569Automated safety check: PassMIT28 days ago
44

Carries out an approved plan for an agtx-managed task: implements the changes, runs tests, commits, writes a summary to .agtx/execute.md and then stops.

fynnfluegge/agtx1.7k—~439Automated safety check: PassApache-2.05 days ago
45

A skill your agent uses when the user asks for a BPMN diagram, swimlane diagram, business process map, or workflow diagram with roles/lanes and phases.

sparklabx/drawio-ai-kit652—~1.7kAutomated safety check: PassMIT5 days ago
46
46.Review ScenarioOfficial

A skill your agent uses when reviewing a conformance PR that adds or changes scenario .ts files for a SEP — before approving, before requesting changes, or as a self-check before opening one.

modelcontextprotocol/conformance129—~983Automated safety check: PassUnknown2 days ago
47

Has the agent build, test and lint a change, re-read the request and check for regressions before it says a coding task is done, then report exactly what it ran.

duckbugio/flock502—~432Automated safety check: PassMIT4 days ago
48

Generates a project-local skill that launches your app, exercises a feature the way a user would and captures evidence, for web, CLI, API or desktop projects.

cursor/plugins10k8 repos~1.5kAutomated safety check: PassNo licenceyesterday

Questions, answered from the data.

What is the best verification before completion skill?

Show Me Your Work Decision Log (official) from cursor/plugins ranks first of the 169 verification before completion skills listed here, with the highest score: its repository has 10k GitHub stars, 9 other GitHub owners carry a copy, its SKILL.md loads about 1.6k tokens and it passes the automated safety check with no findings. Next come Acceptance Evidence for Deliveries and Verification Before Completion.

Which verification before completion skills are official?

8 of the 169 verification before completion skills are official, published by the vendor's own GitHub organization: Show Me Your Work Decision Log, CUTLASS FMHA Incremental Rebuild, Advisor Mode, Review Scenario, Create a Verification Skill and 3 more.

How are these skills ranked?

By Skill Navigator score, which combines the GitHub stars of the skill's repository (shared across that repo's skills and discounted for large collections), how many other GitHub owners carry a copy of the skill, and automated SKILL.md quality checks, minus penalties for safety-check warnings and for each further skill from the same repository. Skills that fail the safety check are not listed.