Verify and Stop
JuliusBrussee/caveman
Prove existing work meets acceptance conditions without expanding scope. Use for validation-only tasks, completion checks, focused gate runs, and last-mile…
Gives a fast single-pass quality verdict on any artifact, with a score from 0 to 1, PASS, REVISE or FAIL, and suggestions for what to change.
$ npx skills add Q00/ouroboros --skill qa -a claude-codeProject install by default; add -g for ~/.claude/skills/.
$ gh skill install Q00/ouroboros qa --agent claude-codeProject scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).
$ git clone --depth 1 https://github.com/Q00/ouroboros.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/qa .claude/skills/qa && rm -rf skills-srcUse ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.
Claude Code skills documentation · loads skills from .claude/skills/
Install the "qa" agent skill from https://github.com/Q00/ouroboros/tree/main/skills/qa into .claude/skills/qa/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "qa", then confirm the skill loads.Claude Code copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$skill-installer install https://github.com/Q00/ouroboros/tree/main/skills/qaType this inside Codex. $skill-installer <name> installs a curated skill from openai/skills. The installer writes to $CODEX_HOME/skills (default ~/.codex/skills). Restart Codex if the skill does not show up.
$ npx skills add Q00/ouroboros --skill qa -a codexProject install goes to .agents/skills/; add -g for ~/.codex/skills/.
$ gh skill install Q00/ouroboros qa --agent codexProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/Q00/ouroboros.git skills-src && mkdir -p .agents/skills && cp -r skills-src/skills/qa .agents/skills/qa && rm -rf skills-srcUse ~/.agents/skills/ instead of .agents/skills for a personal install.
Codex skills documentation · loads skills from .agents/skills/
Install the "qa" agent skill from https://github.com/Q00/ouroboros/tree/main/skills/qa into .agents/skills/qa/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "qa", then confirm the skill loads.Codex copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add Q00/ouroboros --skill qa -a cursorProject install goes to .agents/skills/; add -g for ~/.cursor/skills/.
$ gh skill install Q00/ouroboros qa --agent cursorProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/Q00/ouroboros.git skills-src && mkdir -p .cursor/skills && cp -r skills-src/skills/qa .cursor/skills/qa && rm -rf skills-srcUse ~/.cursor/skills/ instead of .cursor/skills for a personal install.
Cursor skills documentation · loads skills from .cursor/skills/, .agents/skills/, .claude/skills/, .codex/skills/
Install the "qa" agent skill from https://github.com/Q00/ouroboros/tree/main/skills/qa into .cursor/skills/qa/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "qa", then confirm the skill loads.Cursor copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gemini skills install https://github.com/Q00/ouroboros.git --path skills/qa--scope user (default) or --scope workspace; --path is the subfolder of the repo that holds the skill; --consent skips the security confirmation prompt.
$ npx skills add Q00/ouroboros --skill qa -a gemini-cliProject install goes to .agents/skills/; add -g for ~/.gemini/skills/.
$ gh skill install Q00/ouroboros qa --agent gemini-cliProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/Q00/ouroboros.git skills-src && mkdir -p .gemini/skills && cp -r skills-src/skills/qa .gemini/skills/qa && rm -rf skills-srcUse ~/.gemini/skills/ instead of .gemini/skills for a personal install, then run /skills reload.
Gemini CLI skills documentation · loads skills from .gemini/skills/, .agents/skills/
Install the "qa" agent skill from https://github.com/Q00/ouroboros/tree/main/skills/qa into .gemini/skills/qa/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "qa", then confirm the skill loads.Gemini CLI copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gh skill install Q00/ouroboros qaInstalls for Copilot at project scope by default; add --scope user for a personal install. Preview a skill first with gh skill preview. Needs GitHub CLI 2.90.0 or later (public preview).
$ npx skills add Q00/ouroboros --skill qa -a github-copilotProject install goes to .agents/skills/; add -g for ~/.copilot/skills/.
$ git clone --depth 1 https://github.com/Q00/ouroboros.git skills-src && mkdir -p .github/skills && cp -r skills-src/skills/qa .github/skills/qa && rm -rf skills-srcUse ~/.copilot/skills/ instead of .github/skills for a personal install. Commit .github/skills so cloud agent and code review can use it.
GitHub Copilot skills documentation · loads skills from .github/skills/, .claude/skills/, .agents/skills/
Install the "qa" agent skill from https://github.com/Q00/ouroboros/tree/main/skills/qa into .github/skills/qa/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "qa", then confirm the skill loads.GitHub Copilot copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add Q00/ouroboros --skill qa -a opencodeOpenCode documents no install command of its own. Project install goes to .agents/skills/; add -g for ~/.config/opencode/skills/.
$ gh skill install Q00/ouroboros qa --agent opencodeProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/Q00/ouroboros.git skills-src && mkdir -p .opencode/skills && cp -r skills-src/skills/qa .opencode/skills/qa && rm -rf skills-srcUse ~/.config/opencode/skills/ instead of .opencode/skills for a personal install.
OpenCode skills documentation · loads skills from .opencode/skills/, .claude/skills/, .agents/skills/
Install the "qa" agent skill from https://github.com/Q00/ouroboros/tree/main/skills/qa into .opencode/skills/qa/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "qa", then confirm the skill loads.OpenCode copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
qaGives a fast single-pass quality verdict on any artifact, with a score from 0 to 1, PASS, REVISE or FAIL, and suggestions for what to change.
The QA judge works on code, documents, API responses, test output or any custom content, and is the lightweight alternative to a three-stage formal evaluation pipeline in the same toolkit. It first pins down the quality bar, taken from acceptance criteria in a seed YAML, from you, or by asking what good means for the artifact, then rates correctness, completeness, quality, intent alignment and domain-specific points.
The verdict is a score with PASS at 0.80 or above, REVISE from 0.40 to 0.79 and FAIL below 0.40, and each maps to a loop action of done, continue or escalate. The artifact can be a file path, inline text or the most recent execution output. When the Ouroboros QA MCP tool is available the agent calls it, and a fallback mode exists for setups without the MCP server.
4 steps, taken from the first numbered list in SKILL.md.
Read from SKILL.md and the folder at commit 0df5b98. It shows what the files ask for, not the result of running them.
Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.
From allowed-tools in the SKILL.md frontmatter.
No scripts in the folder and no shell commands in SKILL.md.
From the folder's file list and the shell code blocks in SKILL.md.
No URLs in SKILL.md.
From URLs in SKILL.md, links to its own repository left out.
Names no API keys, tokens, secrets or passwords.
From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.
QA Verdict Judge loads about 2.3k tokens when it runs. Until then it costs about 13 tokens; SKILL.md has 945 words of instructions outside code blocks.
Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.
The automated check found no risky patterns in SKILL.md.
Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.
The full file from Q00/ouroboros at commit 0df5b98, republished under its MIT licence (© Q00). 945 words, ~2,251 tokens.
.claude/skills/qa/SKILL.md (or your agent's skills folder).Standalone quality assessment for any artifact — code, documents, API responses, test output, or custom content. Unlike ooo evaluate (3-stage formal verification pipeline), ooo qa is a fast single-pass verdict with actionable suggestions.
ooo qa [file_path | artifact_text]
ooo qa # evaluate recent execution output
/ouroboros:qa [file_path | artifact_text] # plugin modeTrigger keywords: "ooo qa", "qa check", "quality check"
The QA Judge evaluates an artifact against a quality bar and returns a structured verdict:
done (pass), continue (revise), escalate (fail)| Score Range | Verdict | Loop Action |
|---|---|---|
| >= 0.80 | PASS | done |
| 0.40 - 0.79 | REVISE | continue |
| < 0.40 | FAIL | escalate |
When the user invokes this skill:
This skill works in two modes. Determine which one before attempting any tool calls:
MCP mode — If the QA MCP tool is available (already exposed, or loadable via discovery), use it:
tool discovery query: "+ouroboros qa"If found (typically named mcp__plugin_ouroboros_ouroboros__ouroboros_qa), proceed with QA Steps below.
Fallback mode — Only if the QA MCP tool is genuinely absent (no Ouroboros MCP server) skip to the Fallback section; an empty discovery result for an already-exposed tool is expected — call it directly rather than falling back. This skill is designed to work without MCP setup.
Determine the artifact to evaluate:
Determine the quality bar:
Determine artifact type:
code — source code filestest_output — test results, CI outputdocument — specs, docs, READMEsapi_response — API responses, JSON payloadsscreenshot — visual artifactscustom — anything else3.5. Acting verification fan-out — probe in parallel, then judge (do not skip for behaviour-bearing artifacts): A text judge can be fooled by a hopeful log line. When the artifact actually does something (code, an app, an API, a UI), fan out empirical probes using the host's native parallel sub-agent primitive — one probe sub-agent per acting modality the runtime actually exposes, all spawned in the same message so they run concurrently:
Bash/shell): run the command / start the app / run
the declared smoke commands with bounded timeouts; capture exit codes and
real output.misleading_output (claimed success vs. real effect), hung_command
(bounded timeout?), malformed_input, stale_state, dirty_worktree.
Skip a modality only when its tools are absent or the artifact type makes
it meaningless — and say which modalities were skipped and why. Await all probes, then pass the merged evidence into the judge as
reference (prefer observed behaviour over source text as the artifact
when they disagree). Empirical evidence outranks the judge: if the
judge scores PASS but any probe observed the behaviour failing, present
the verdict as REVISE/FAIL on that evidence and say so explicitly — a
score contradicted by observation is not a pass. If no acting tools are
available at all, judge on the text alone but flag that behaviour was not
observed.
Call the ouroboros_qa MCP tool:
Tool: ouroboros_qa
Arguments:
artifact: <the content to evaluate>
quality_bar: <what 'pass' means>
artifact_type: "code" (or other type)
reference: <observed-behaviour evidence from step 3.5, plus any reference>
pass_threshold: 0.80 (adjustable)
seed_content: <seed YAML if available>Present results clearly:
Next: Your artifact meets the quality bar. Proceed with confidence.Next: Address the suggestions above, then run ooo qa again to re-check.Next: Fundamental issues detected. Consider ooo interview to re-examine requirements, or ooo unstuck to challenge assumptions.For iterative usage, track the qa_session_id and iteration_history from the response meta:
qa_session_id and iteration_entry in metaqa_session_id and accumulated iteration_historypass or failIn fallback mode, generate a qa-<uuid4_short> session ID on the first run and maintain iteration count in conversation context to preserve the same iterative contract.
If the MCP server is not available, adopt the ouroboros:qa-judge agent role directly:
<project-root>/src/ouroboros/agents/qa-judge.md
(This is the same prompt used by the MCP QA tool, ensuring consistent verdicts.)QA Verdict [Iteration N]
========================
Session: qa-<id>
Score: X.XX / 1.00 [PASS/REVISE/FAIL]
Verdict: pass/revise/fail
Threshold: 0.80
Dimensions:
Correctness: X.XX
Completeness: X.XX
Quality: X.XX
Intent Alignment: X.XX
Domain-Specific: X.XX
Differences:
- <specific difference>
Suggestions:
- <actionable fix>
Reasoning: <1-3 sentence summary>
Loop Action: done/continue/escalateUser: ooo qa src/main.py
QA Verdict [Iteration 1]
============================================================
Session: qa-a1b2c3d4
Score: 0.72 / 1.00 [REVISE]
Verdict: revise
Threshold: 0.80
Dimensions:
Correctness: 0.85
Completeness: 0.60
Quality: 0.75
Intent Alignment: 0.80
Domain-Specific: 0.60
Differences:
- Missing error handling for network timeout in fetch_data()
- No input validation on user_id parameter
- Type hints missing on 3 public functions
Suggestions:
- Add try/except with TimeoutError in fetch_data() (line 42)
- Add isinstance check for user_id at function entry
- Add return type annotations to get_user(), fetch_data(), process_result()
Reasoning: Core logic is correct but lacks defensive programming
patterns expected for production code.
Loop Action: continue
Next: Address the suggestions above, then run `ooo qa` again to re-check.Your final response MUST end with exactly one breadcrumb footer line:
◆ <current state> → next: <recommended action>Derive <current state> from live session state via ouroboros_session_status when that MCP projection is available; otherwise derive it from this skill's actual outcome. Never use a linear Step N of M footer because Ouroboros is an evolutionary loop. When the next action is genuinely a choice, list 2-3 honest options in the next: clause. The breadcrumb line must be the last line of the response.
© Q00, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file
Just SKILL.md in skills/qa of Q00/ouroboros.
Open the folder on GitHubat commit 0df5b98
QA Verdict Judge next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.
| Skill | Stars | Used in | Tokens | Auto-check | Licence | Repo updated |
|---|---|---|---|---|---|---|
| QA Verdict Judge this skillQ00/ouroboros | 6.2k | — | ~2.3k | Automated safety check: Pass | MIT | |
| Verify and StopJuliusBrussee/caveman | 110k | 1 repos | ~176 | Automated safety check: Pass | Apache-2.0 | |
| Iteration Progress Auditprime-radiant-inc/iterative-development | 181 | — | ~1.1k | Automated safety check: Pass | Apache-2.0 | |
| Verification Before CompletionjnMetaCode/superpowers-zh | 8.3k | — | ~443 | Automated safety check: Pass | MIT | |
| Verify Before CompletionYeachan-Heo/oh-my-claudecode | 40k | — | ~277 | Automated safety check: Pass | MIT | |
| Auto Review LoopCurryTang/Amadeus | 176 | — | ~4.3k | Automated safety check: Notes | None |
JuliusBrussee/caveman
Prove existing work meets acceptance conditions without expanding scope. Use for validation-only tasks, completion checks, focused gate runs, and last-mile…
prime-radiant-inc/iterative-development
Checks the quality of behavior evidence after each iteration in three tiers, using two auditor subagents in parallel to review the same work and find gaps.
jnMetaCode/superpowers-zh
Chinese-language rule that bars an agent from claiming work is done, fixed or passing until it has run a verification command and read the output.
Yeachan-Heo/oh-my-claudecode
Has the agent prove that a feature, fix or refactor works, using existing tests first, then narrow commands and manual checks, and report only what was actually verified.
CurryTang/Amadeus
Autonomous multi-round research review loop. An agent skill from CurryTang/Amadeus.
JasonColapietro/suede-creator-skills
Checks a Suede AI MCP server release against a live process: the full JSON-RPC lifecycle, schemas, annotations, malformed input, catalog agreement and install docs.
Q00/ouroboros
Triages and works through GitHub issues and pull requests in the Q00/ouroboros repo as a maintainer, within a stated review boundary and clear limits on what it may change.
Q00/ouroboros
Runs a guided product-manager interview that classifies each question automatically and produces a Product Requirements Document.
Q00/ouroboros
Scans a directory for existing git repositories and worktrees, then registers and manages which ones serve as default context during interviews.
Q00/ouroboros
Scores an agent's finished work with a three-stage pipeline: free mechanical checks, an advisory semantic review, and an optional multi-model consensus vote.
Q00/ouroboros
Starts, monitors or rewinds an evolutionary development loop that refines an ontology and acceptance criteria generation by generation until it converges, using the Ouroboros MCP tools.
Q00/ouroboros
Opens or drives the Ouroboros settings GUI, picking a browser, TUI or chat-based approach depending on whether the user can reach a browser window.
Works with
Categories
Gives a fast single-pass quality verdict on any artifact, with a score from 0 to 1, PASS, REVISE or FAIL, and suggestions for what to change. The QA judge works on code, documents, API responses, test output or any custom content, and is the lightweight alternative to a three-stage formal evaluation pipeline in the same toolkit. It first pins down the quality bar, taken from acceptance criteria in a seed YAML, from you, or by asking what good means for the artifact, then rates correctness, completeness, quality, intent alignment and domain-specific points.
QA Verdict Judge fits situations like: getting a quick pass or revise verdict on a draft, patch or document; checking an artifact against explicit acceptance criteria; deciding whether an agent loop should finish, continue or escalate.
Run `npx skills add Q00/ouroboros --skill qa -a claude-code`. Or copy the skill folder (skills/qa in Q00/ouroboros) into .claude/skills/qa in your project. Claude Code loads it when a task matches its description.
Run `npx skills add Q00/ouroboros --skill qa -a codex`. Or copy the skill folder (skills/qa in Q00/ouroboros) into .agents/skills/qa in your project. Codex loads it when a task matches its description.
Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add Q00/ouroboros --skill qa -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/qa, .gemini/skills/qa, .github/skills/qa and .opencode/skills/qa in your project.
SKILL.md names no scripts, command-line tools or credentials: QA Verdict Judge is instructions for the agent only. Our summary lists: The Ouroboros QA MCP tool, or the skill's fallback mode without it.
SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.
Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.
QA Verdict Judge is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.
About 2.3k tokens (SKILL.md is roughly 9k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.
Skills that share tags, products or a category with QA Verdict Judge: Verify and Stop (JuliusBrussee/caveman, 110k stars), Iteration Progress Audit (prime-radiant-inc/iterative-development, 181 stars), Verification Before Completion (jnMetaCode/superpowers-zh, 8.3k stars) and Verify Before Completion (Yeachan-Heo/oh-my-claudecode, 40k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.
Q00 (a GitHub user) maintains it in Q00/ouroboros, which has 6,189 GitHub stars. The repository holds 23 skills in this directory. The repository was last updated on October 6, 2026.
Source: Q00/ouroboros on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.