Topic · Agent Workflows
Best verification before completion skills for Claude Code, Codex and other agents.
- skills
- 169
- official
- 8
Verification before completion skills, ranked
Ranked by score. Sort bymost stars,trending,newest,recently updated
| # | Skill | Repository | Stars | Used in | Tokens | Auto-check | Licence | Updated |
|---|---|---|---|---|---|---|---|---|
| 1 | Keeps a TSV decision log for long or unattended agent runs, one row per decision with what, why, evidence and result, so a reviewer can check the work later. | cursor/ | 10k | 9 repos | ~1.6k | Automated safety check: Pass | No licence | yesterday |
| 2 | Verifies a delivery end to end by driving the real product on a CLI, web, desktop or iOS Simulator surface, capturing evidence and publishing a round with the lh CLI. | lobehub/ | 83k | — | ~9.7k | Automated safety check: Pass | Apache-2.0 | today |
| 3 | A skill your agent uses when about to claim work is complete, fixed, or passing, before committing or creating PRs - requires running verification commands and confirming output before making any… | farm-fe/ | 5.6k | 45 repos | ~1k | Automated safety check: Pass | MIT | 15 days ago |
| 4 | Keeps an agent focused on the requested work by applying a five-step ladder that checks for direct solutions, real gaps and speculative defenses before adding anything. | lennney/ | 2.5k | 1 repo | ~2k | Automated safety check: Pass | MIT | 2 days ago |
| 5 | Adds short, pointed workplace reminders drawn from two Chinese essays on corporate culture, nudging the agent to prove results with evidence instead of polished reports. | tanweai/ | 20k | — | ~556 | Automated safety check: Pass | MIT | 28 days ago |
| 6 | Proves a Codewhale change in the real product: a stamped release build, an atomic local install, fresh-shell verification and manual QA that automated gates cannot cover. | codewhale-hq/ | 41k | — | ~1.3k | Automated safety check: Pass | MIT | today |
| 7 | Explains why editing CUTLASS fused-MHA headers in ONNX Runtime can leave stale CUDA kernels after an incremental build, and how to force and verify a real rebuild. | microsoft/ | 22k | — | ~1.3k | Automated safety check: Pass | MIT | today |
| 8 | Records an annotated screen recording of the agent testing an app hands-on, then posts the video and a results summary to the PR and tracker issue. | michaelshimeles/ | 1.3k | 1 repo | ~3.9k | Automated safety check: Pass | No licence | 3 days ago |
| 9 | Read and write email through the Atomic Mail from an AI agent. | Atomic-Mail/ | 266 | 1 repo | ~2k | Automated safety check: Pass | MIT | 10 days ago |
| 10 | Pushes an agent to keep verifying and changing approach after repeated failures, using a diagnosis line, evidence-based completion and confirmation before risky edits. | tanweai/ | 20k | — | ~502 | Automated safety check: Pass | MIT | 28 days ago |
| 11 | The binding directive for ultrawork mode: evidence-first delivery, a light or heavy process tier, and constant use of memory during the work. | code-yeongyu/ | 3.7k | — | ~6.9k | Automated safety check: Pass | MIT | yesterday |
| 12 | Delivers a change in thin vertical slices, each implemented, tested, verified and committed before the next, using vertical, contract-first or risk-first slicing. | addyosmani/ | 102k | 1 repo | ~2.3k | Automated safety check: Pass | MIT | 4 days ago |
| 13 | 13.PUA Loop Runs an unattended iterate-until-verified loop in which a user-set verify command, not the agent's own claim, decides when the task is finished. | tanweai/ | 20k | — | ~1.1k | Automated safety check: Pass | MIT | 28 days ago |
| 14 | Executes a written implementation plan task by task in the current session, with a progress ledger, test-first gates and one fresh-context review at the end. | jnMetaCode/ | 8.3k | — | ~2.5k | Automated safety check: Pass | MIT | 3 days ago |
| 15 | Runs the Antigravity CLI headlessly with the agy command to analyze a repository or carry out a coding task, then checks its JSON result and the diff. | CherryHQ/ | 52k | — | ~531 | Automated safety check: Pass | AGPL-3.0 | today |
| 16 | Confirms that a change to PlotJuggler 4 really works in the running app by proving the rebuild, launching with real data and measuring the result. | PlotJuggler/ | 6.2k | — | ~956 | Automated safety check: Pass | MPL-2.0 | 6 days ago |
| 17 | 17.Task Creator SkillsBench task authoring — walk a contributor from idea to submission-ready task following CONTRIBUTING.md and the task-implementation rubric. | benchflow-ai/ | 353 | — | ~4.5k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 18 | Researches code with evidence: traces callers, imports and cross-repo links, diagnoses failures and reports findings with exact file and line references and a confidence label. | bgauryy/ | 946 | — | ~1.5k | Automated safety check: Pass | MIT | 4 days ago |
| 19 | A skill your agent uses when an AI agent wants to participate in the Haidian Centennial Jing-Zhang AI Innovation Belt open call, follow changing materials and community discussion, generate or… | open-city-ai/ | 415 | — | ~9.4k | Automated safety check: Pass | No licence | 1 mo ago |
| 20 | A skill your agent uses when about to claim a result, effect, significance, or that an analysis reproduces, before reporting or writing it up - requires running the analysis fresh and reading the… | K-Dense-AI/ | 348 | 1 repo | ~1.5k | Automated safety check: Pass | Unknown | 24 days ago |
| 21 | Decomposes a complex feature into tasks, dispatches parallel specialist agents with durable state, and supervises verification, QA review and retries. | first-fluke/ | 1.3k | — | ~4.1k | Automated safety check: Pass | MIT | yesterday |
| 22 | 22.Drawio AWS A skill your agent uses when the user asks for an AWS architecture diagram — VPC/networking, event-driven, landing zone, multi-AZ, serverless pipeline, or any diagram built with AWS service icons. | sparklabx/ | 652 | — | ~1.6k | Automated safety check: Pass | MIT | 5 days ago |
| 23 | 23.Docs Guard Checks generated or edited documentation against the source code, flagging invented symbols, outdated samples and unverifiable claims before publishing. | amElnagdy/ | 1.3k | — | ~2.1k | Automated safety check: Pass | MIT | 3 mo ago |
| 24 | Adds a second, stronger model that the main agent consults before major decisions, when stuck and before finishing, controlled by /advisor commands. | cursor/ | 10k | — | ~2.6k | Automated safety check: Notes | No licence | yesterday |
| 25 | A gate before merge: verify runs the real app against the spec, and review has a different model do a senior code review, without editing code. | jsmastery-pro/ | 1.4k | — | ~1.1k | Automated safety check: Notes | MIT | 1 mo ago |
| 26 | Coordinates and recovers multi-stage Light research projects from a single passport file, with checkpoints, stale-work tracking and rerouting only when you approve. | Light0305/ | 640 | — | ~3.8k | Automated safety check: Pass | MIT | 3 mo ago |
| 27 | Runs an evidence-bound audit, daily or weekly review through the requirement-ledger CLI, binding scope and authority first and blocking the final report until a handoff check passes. | adand-91/ | 125 | — | ~1.5k | Automated safety check: Pass | MIT | 21 days ago |
| 28 | Turns a one-line idea into a task brief of up to 4000 characters that an autonomous agent can run through /goal, with measured baselines and cheat-resistant acceptance checks. | KKKKhazix/ | 21k | — | ~723 | Automated safety check: Pass | MIT | 6 days ago |
| 29 | Pushes an agent that keeps failing or gives up to exhaust every option, using harsh corporate-pressure wording, a diagnosis line and a proactivity checklist. | tanweai/ | 20k | — | ~3.3k | Automated safety check: Pass | MIT | 28 days ago |
| 30 | 30.Drawio Azure A skill your agent uses when the user asks for an Azure architecture diagram — VNet/networking, App Service, AKS, landing zone, multi-region, or any diagram built with Azure service icons. | sparklabx/ | 652 | — | ~1.6k | Automated safety check: Pass | MIT | 5 days ago |
| 31 | Applies a 12-stage verified workflow, from research to deploy, to non-trivial coding tasks, scaled to lightweight, standard or full mode by task size. | artemiimillier/ | 153 | — | ~3.5k | Automated safety check: Pass | MIT | 6 mo ago |
| 32 | Proves the Puppetmaster Jev transition behaves as specified: unset never uses the network, opt-in may skip the conflict auditor, and later versions only observe. | professorpalmer/ | 467 | — | ~525 | Automated safety check: Pass | MIT | today |
| 33 | Final checklist before handing work back in the herdr-auto-title repo: review the diff, run make check, apply the comment and AGENTS.md rules, then report. | kryptamine/ | 237 | — | ~605 | Automated safety check: Pass | MIT | yesterday |
| 34 | Run the infrastructure health check and fix anything that fails | diet103/ | 10k | — | ~363 | Automated safety check: Pass | MIT | 2 mo ago |
| 35 | Read-only detector that compares SPEC.md with the code and reports invariant, interface and task drift grouped by severity, without changing anything. | JuliusBrussee/ | 1.1k | — | ~666 | Automated safety check: Pass | MIT | 1 mo ago |
| 36 | Scores an agent's finished work with a three-stage pipeline: free mechanical checks, an advisory semantic review, and an optional multi-model consensus vote. | Q00/ | 6.2k | — | ~2.2k | Automated safety check: Pass | MIT | yesterday |
| 37 | Human handoff, Linear branch names, reviewable (non-draft) PRs, the cursor GitHub label, Claude, Greptile, and Codex review comments, preview test steps, proof of work posted on the GitHub PR, and… | langfuse/ | 35k | — | ~1.7k | Automated safety check: Pass | Unknown | today |
| 38 | 38.Fable Mode A work-discipline protocol that makes Opus 4.8 (or any non-frontier model) operate at Fable-5-grade quality. | cozytab/ | 106 | — | ~4.8k | Automated safety check: Pass | MIT | 2 mo ago |
| 39 | A skill your agent uses when the user wants to set up / scaffold / install a file-based multi-agent orchestration system in a folder. | netwaif/ | 106 | — | ~629 | Automated safety check: Pass | MIT | 1 mo ago |
| 40 | Applies Procoder's senior-developer discipline in a repository: run the commit gate, format through the binary and work through specs, plans and todos. | azrtydxb/ | 211 | — | ~3.7k | Automated safety check: Pass | Apache-2.0 | 9 days ago |
| 41 | A required checklist for after code changes: validate each changed layer, check that docs and other layers stay in sync, and test before committing or opening a PR. | ZeroDeng01/ | 1.7k | — | ~4.4k | Automated safety check: Pass | MIT | 2 days ago |
| 42 | Runs a math modeling contest pipeline for CUMCM and MCM/ICM entries in Codex CLI, with git checkpoints, verified solver code and human review at each stage. | RealSeaberry/ | 258 | — | ~1.6k | Automated safety check: Pass | MIT | 27 days ago |
| 43 | Pushes an agent that keeps failing, gives up or claims unverified success into a diagnosis, evidence and verification loop, with a Pi extension for persistent mode. | tanweai/ | 20k | — | ~569 | Automated safety check: Pass | MIT | 28 days ago |
| 44 | Carries out an approved plan for an agtx-managed task: implements the changes, runs tests, commits, writes a summary to .agtx/execute.md and then stops. | fynnfluegge/ | 1.7k | — | ~439 | Automated safety check: Pass | Apache-2.0 | 5 days ago |
| 45 | 45.Drawio Bpmn A skill your agent uses when the user asks for a BPMN diagram, swimlane diagram, business process map, or workflow diagram with roles/lanes and phases. | sparklabx/ | 652 | — | ~1.7k | Automated safety check: Pass | MIT | 5 days ago |
| 46 | A skill your agent uses when reviewing a conformance PR that adds or changes scenario .ts files for a SEP — before approving, before requesting changes, or as a self-check before opening one. | modelcontextprotocol/ | 129 | — | ~983 | Automated safety check: Pass | Unknown | 2 days ago |
| 47 | Has the agent build, test and lint a change, re-read the request and check for regressions before it says a coding task is done, then report exactly what it ran. | duckbugio/ | 502 | — | ~432 | Automated safety check: Pass | MIT | 4 days ago |
| 48 | Generates a project-local skill that launches your app, exercises a feature the way a user would and captures evidence, for web, CLI, API or desktop projects. | cursor/ | 10k | 8 repos | ~1.5k | Automated safety check: Pass | No licence | yesterday |
Questions, answered from the data.
What is the best verification before completion skill?
Show Me Your Work Decision Log (official) from cursor/plugins ranks first of the 169 verification before completion skills listed here, with the highest score: its repository has 10k GitHub stars, 9 other GitHub owners carry a copy, its SKILL.md loads about 1.6k tokens and it passes the automated safety check with no findings. Next come Acceptance Evidence for Deliveries and Verification Before Completion.
Which verification before completion skills are official?
8 of the 169 verification before completion skills are official, published by the vendor's own GitHub organization: Show Me Your Work Decision Log, CUTLASS FMHA Incremental Rebuild, Advisor Mode, Review Scenario, Create a Verification Skill and 3 more.
How are these skills ranked?
By Skill Navigator score, which combines the GitHub stars of the skill's repository (shared across that repo's skills and discounted for large collections), how many other GitHub owners carry a copy of the skill, and automated SKILL.md quality checks, minus penalties for safety-check warnings and for each further skill from the same repository. Skills that fail the safety check are not listed.
Explore related skills
Category
More topics in Agent Workflows
- MCP servers2,017
- Subagents1,015
- Agent instruction files737
- Agent memory507
- Multi-agent orchestration504
- Skill authoring502
- Planning495
- Brainstorming467
- Hooks and plugins363
- Autonomous loops325
- Context engineering297
- Session handoff262
- Requirements gathering234
- Task breakdown232
- Skill management205
- Human-in-the-loop approvals203
- Codebase knowledge for agents186
- Agent evaluation and testing131