DeerFlow Smoke Test
bytedance/deer-flow
Walks through an end-to-end smoke test of a DeerFlow deployment: pull the latest code, deploy with Docker or locally, verify services, run health checks and write a report.
This skill should be used when the user asks to "verify a fix", "reproduce failure", "diagnose issue", "check BEFORE/AFTER state", "VF task", "reality check", "check test quality", "mock-only…
$ npx skills add tzachbon/smart-ralph --skill reality-verification -a claude-codeProject install by default; add -g for ~/.claude/skills/.
$ gh skill install tzachbon/smart-ralph reality-verification --agent claude-codeProject scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).
$ git clone --depth 1 https://github.com/tzachbon/smart-ralph.git skills-src && mkdir -p .claude/skills && cp -r skills-src/plugins/ralph-specum/skills/reality-verification .claude/skills/reality-verification && rm -rf skills-srcUse ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.
Claude Code skills documentation · loads skills from .claude/skills/
Install the "reality-verification" agent skill from https://github.com/tzachbon/smart-ralph/tree/main/plugins/ralph-specum/skills/reality-verification into .claude/skills/reality-verification/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "reality-verification", then confirm the skill loads.Claude Code copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$skill-installer install https://github.com/tzachbon/smart-ralph/tree/main/plugins/ralph-specum/skills/reality-verificationType this inside Codex. $skill-installer <name> installs a curated skill from openai/skills. The installer writes to $CODEX_HOME/skills (default ~/.codex/skills). Restart Codex if the skill does not show up.
$ npx skills add tzachbon/smart-ralph --skill reality-verification -a codexProject install goes to .agents/skills/; add -g for ~/.codex/skills/.
$ gh skill install tzachbon/smart-ralph reality-verification --agent codexProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/tzachbon/smart-ralph.git skills-src && mkdir -p .agents/skills && cp -r skills-src/plugins/ralph-specum/skills/reality-verification .agents/skills/reality-verification && rm -rf skills-srcUse ~/.agents/skills/ instead of .agents/skills for a personal install.
Codex skills documentation · loads skills from .agents/skills/
Install the "reality-verification" agent skill from https://github.com/tzachbon/smart-ralph/tree/main/plugins/ralph-specum/skills/reality-verification into .agents/skills/reality-verification/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "reality-verification", then confirm the skill loads.Codex copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add tzachbon/smart-ralph --skill reality-verification -a cursorProject install goes to .agents/skills/; add -g for ~/.cursor/skills/.
$ gh skill install tzachbon/smart-ralph reality-verification --agent cursorProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/tzachbon/smart-ralph.git skills-src && mkdir -p .cursor/skills && cp -r skills-src/plugins/ralph-specum/skills/reality-verification .cursor/skills/reality-verification && rm -rf skills-srcUse ~/.cursor/skills/ instead of .cursor/skills for a personal install.
Cursor skills documentation · loads skills from .cursor/skills/, .agents/skills/, .claude/skills/, .codex/skills/
Install the "reality-verification" agent skill from https://github.com/tzachbon/smart-ralph/tree/main/plugins/ralph-specum/skills/reality-verification into .cursor/skills/reality-verification/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "reality-verification", then confirm the skill loads.Cursor copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gemini skills install https://github.com/tzachbon/smart-ralph.git --path plugins/ralph-specum/skills/reality-verification--scope user (default) or --scope workspace; --path is the subfolder of the repo that holds the skill; --consent skips the security confirmation prompt.
$ npx skills add tzachbon/smart-ralph --skill reality-verification -a gemini-cliProject install goes to .agents/skills/; add -g for ~/.gemini/skills/.
$ gh skill install tzachbon/smart-ralph reality-verification --agent gemini-cliProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/tzachbon/smart-ralph.git skills-src && mkdir -p .gemini/skills && cp -r skills-src/plugins/ralph-specum/skills/reality-verification .gemini/skills/reality-verification && rm -rf skills-srcUse ~/.gemini/skills/ instead of .gemini/skills for a personal install, then run /skills reload.
Gemini CLI skills documentation · loads skills from .gemini/skills/, .agents/skills/
Install the "reality-verification" agent skill from https://github.com/tzachbon/smart-ralph/tree/main/plugins/ralph-specum/skills/reality-verification into .gemini/skills/reality-verification/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "reality-verification", then confirm the skill loads.Gemini CLI copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gh skill install tzachbon/smart-ralph reality-verificationInstalls for Copilot at project scope by default; add --scope user for a personal install. Preview a skill first with gh skill preview. Needs GitHub CLI 2.90.0 or later (public preview).
$ npx skills add tzachbon/smart-ralph --skill reality-verification -a github-copilotProject install goes to .agents/skills/; add -g for ~/.copilot/skills/.
$ git clone --depth 1 https://github.com/tzachbon/smart-ralph.git skills-src && mkdir -p .github/skills && cp -r skills-src/plugins/ralph-specum/skills/reality-verification .github/skills/reality-verification && rm -rf skills-srcUse ~/.copilot/skills/ instead of .github/skills for a personal install. Commit .github/skills so cloud agent and code review can use it.
GitHub Copilot skills documentation · loads skills from .github/skills/, .claude/skills/, .agents/skills/
Install the "reality-verification" agent skill from https://github.com/tzachbon/smart-ralph/tree/main/plugins/ralph-specum/skills/reality-verification into .github/skills/reality-verification/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "reality-verification", then confirm the skill loads.GitHub Copilot copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add tzachbon/smart-ralph --skill reality-verification -a opencodeOpenCode documents no install command of its own. Project install goes to .agents/skills/; add -g for ~/.config/opencode/skills/.
$ gh skill install tzachbon/smart-ralph reality-verification --agent opencodeProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/tzachbon/smart-ralph.git skills-src && mkdir -p .opencode/skills && cp -r skills-src/plugins/ralph-specum/skills/reality-verification .opencode/skills/reality-verification && rm -rf skills-srcUse ~/.config/opencode/skills/ instead of .opencode/skills for a personal install.
OpenCode skills documentation · loads skills from .opencode/skills/, .claude/skills/, .agents/skills/
Install the "reality-verification" agent skill from https://github.com/tzachbon/smart-ralph/tree/main/plugins/ralph-specum/skills/reality-verification into .opencode/skills/reality-verification/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "reality-verification", then confirm the skill loads.OpenCode copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
reality-verificationThis skill should be used when the user asks to "verify a fix", "reproduce failure", "diagnose issue", "check BEFORE/AFTER state", "VF task", "reality check", "check test quality", "mock-only…
Reality Verification is an agent skill from tzachbon/smart-ralph. This skill should be used when the user asks to "verify a fix", "reproduce failure", "diagnose issue", "check BEFORE/AFTER state", "VF task", "reality check", "check test quality", "mock-only tests", or needs guidance on verifying fixes by reproducing failures before and after implementation, or detecting mock-heavy test anti-patterns.
Its SKILL.md is about 880 tokens, which your agent loads only when the skill is triggered. The skill folder holds 3 other files, including reference files (for example `references/goal-detection-patterns.md` and `references/mock-quality-checks.md`).
It sits in Testing & QA. It works with pnpm. The repository describes itself as: Spec-driven development with smart compaction. Claude Code plugin combining Ralph Wiggum loop with structured specification workflow. The licence is MIT.
Read from SKILL.md and the folder at commit ac7251a. It shows what the files ask for, not the result of running them.
Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.
From allowed-tools in the SKILL.md frontmatter.
Shell commands in SKILL.md call:
pnpmghtscFrom the folder's file list and the shell code blocks in SKILL.md.
No URLs in SKILL.md. Its commands use pnpm and gh, which can reach the network depending on how they are called.
From URLs in SKILL.md, links to its own repository left out.
Names no API keys, tokens, secrets or passwords.
From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.
Reality Verification loads about 877 tokens when it runs, and up to ~2.3k if it reads all its reference files. Until then it costs about 90 tokens; SKILL.md has 272 words of instructions outside code blocks.
Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.
The automated check found no risky patterns in SKILL.md.
Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.
The full file from tzachbon/smart-ralph at commit ac7251a, republished under its MIT licence (© tzachbon). 272 words, ~877 tokens.
.claude/skills/reality-verification/SKILL.md (or your agent's skills folder). This skill also uses 2 other files; get the full folder from GitHub.For fix goals: reproduce the failure BEFORE work, verify resolution AFTER.
Classify user goals to determine if diagnosis is needed. See references/goal-detection-patterns.md for detailed patterns.
Quick reference:
| Goal Keywords | Reproduction Command |
|---|---|
| CI, pipeline | gh run view --log-failed |
| test, tests | project test command |
| type, typescript | pnpm check-types or tsc --noEmit |
| lint | pnpm lint |
| build | pnpm build |
| E2E, UI | Playwright MCP browser tools |
| API, endpoint | WebFetch tool |
For E2E/deployment verification, use MCP tools (Playwright MCP browser tools for UI, WebFetch tool for APIs).
Document in .progress.md under ## Reality Check (BEFORE):
## Reality Check (BEFORE)
**Goal type**: Fix
**Reproduction command**: `pnpm test`
**Failure observed**: Yes
**Output**:FAIL src/auth.test.ts Expected: 200 Received: 401
**Timestamp**: 2026-01-16T10:30:00ZDocument in .progress.md under ## Reality Check (AFTER):
## Reality Check (AFTER)
**Command**: `pnpm test`
**Result**: PASS
**Output**:PASS src/auth.test.ts All tests passed
**Comparison**: BEFORE failed with 401, AFTER passes
**Verified**: Issue resolvedAdd as task 4.3 (after PR creation) for fix-type specs:
- [ ] 4.3 VF: Verify original issue resolved
- **Do**:
1. Read BEFORE state from .progress.md
2. Re-run reproduction command: `<command>`
3. Compare output with BEFORE state
4. Document AFTER state in .progress.md
- **Verify**: `grep -q "Verified: Issue resolved" ./specs/<name>/.progress.md`
- **Done when**: AFTER shows issue resolved, documented in .progress.md
- **Commit**: `chore(<name>): verify fix resolves original issue`When verifying test-related fixes, check for mock-only test anti-patterns. See references/mock-quality-checks.md for detailed patterns.
Quick reference red flags:
| Without | With |
|---|---|
| "Fix CI" spec completes but CI still red | CI verified green before merge |
| Tests "fixed" but original failure unknown | Before/after comparison proves fix |
| Silent regressions | Explicit failure reproduction |
| Manual verification required | Automated verification in workflow |
| Tests pass but only test mocks | Tests verify real behavior, not mock behavior |
| False sense of security from green tests | Confidence that tests catch real bugs |
© tzachbon, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file
SKILL.md and 2 other files (references) in plugins/ralph-specum/skills/reality-verification of tzachbon/smart-ralph.
Open the folder on GitHubat commit ac7251a
Reality Verification next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.
| Skill | Stars | Used in | Tokens | Auto-check | Licence | Repo updated |
|---|---|---|---|---|---|---|
| Reality Verification this skilltzachbon/smart-ralph | 558 | — | ~877 | Automated safety check: Pass | MIT | |
| DeerFlow Smoke Testbytedance/deer-flow | 84k | — | ~2.5k | Automated safety check: Notes | MIT | |
| OpenWork Desktop CDP Driverdifferent-ai/openwork | 24k | — | ~465 | Automated safety check: Pass | Custom licence | |
| E2Egronxb/hot-updater | 1.8k | — | ~1.6k | Automated safety check: Pass | Custom licence | |
| Cloud Agents Starterscalar/scalar | 16k | — | ~1.4k | Automated safety check: Pass | MIT | |
| Desktop Dual Repo TestEynzof/Hermes-CN-Desktop | 1.8k | — | ~1.8k | Automated safety check: Pass | Custom licence |
bytedance/deer-flow
Walks through an end-to-end smoke test of a DeerFlow deployment: pull the latest code, deploy with Docker or locally, verify services, run health checks and write a report.
different-ai/openwork
Drives a running OpenWork desktop window over CDP from the shell to evaluate JS, take screenshots, start sessions and send prompts for hand checks.
gronxb/hot-updater
Run end-to-end OTA verification for examples/v0.85.0 with agent-device.
scalar/scalar
Minimal starter runbook for cloud agents to install dependencies, run packages, execute tests, and troubleshoot the Scalar monorepo quickly.
Eynzof/Hermes-CN-Desktop
A skill your agent uses when starting or verifying Hermes Agent CN Desktop with the latest Hermes-CN-Desktop and Hermes-CN-Core branches — dev smoke test (pnpm tauri:dev), packaged beta/release…
module-federation/core
Run this repository's local CI parity commands and pnpm run ci:local jobs.
tzachbon/smart-ralph
This skill should be used when generating spec artifacts (research.md, requirements.md, design.md, tasks.md), formatting agent output, structuring phase results, or when any Ralph agent needs…
tzachbon/smart-ralph
This skill should be used when a Ralph phase must identify critical user decisions, run a layered grill, persist partial answers, obtain explicit approval, or resume an interrupted phase interview…
tzachbon/smart-ralph
This skill should be used when the user asks to "build a feature", "create a spec", "start spec-driven development", "run research phase", "generate requirements", "create design", "plan tasks"…
tzachbon/smart-ralph
This skill should be used only when the user explicitly asks to use $ralph-specum-design, or explicitly asks Ralph Specum in Codex to run the design phase.
tzachbon/smart-ralph
This skill should be used only when the user explicitly asks to use $ralph-specum-tasks, or explicitly asks Ralph Specum in Codex to run the tasks phase.
tzachbon/smart-ralph
This skill should be used when the user asks about "ralph arguments", "quick mode", "commit spec", "max iterations", "ralph state file", "prototype overlay", "execution modes", "ralph loop"…
Works with
Categories
This skill should be used when the user asks to "verify a fix", "reproduce failure", "diagnose issue", "check BEFORE/AFTER state", "VF task", "reality check", "check test quality", "mock-only…. Reality Verification is an agent skill from tzachbon/smart-ralph. This skill should be used when the user asks to "verify a fix", "reproduce failure", "diagnose issue", "check BEFORE/AFTER state", "VF task", "reality check", "check test quality", "mock-only tests", or needs guidance on verifying fixes by reproducing failures before and after implementation, or detecting mock-heavy test anti-patterns.
Reality Verification fits situations like: asks to verify a fix; reproduce failure; check BEFORE/AFTER state; check test quality.
Run `npx skills add tzachbon/smart-ralph --skill reality-verification -a claude-code`. Or copy the skill folder (plugins/ralph-specum/skills/reality-verification in tzachbon/smart-ralph) into .claude/skills/reality-verification in your project. Claude Code loads it when a task matches its description.
Run `npx skills add tzachbon/smart-ralph --skill reality-verification -a codex`. Or copy the skill folder (plugins/ralph-specum/skills/reality-verification in tzachbon/smart-ralph) into .agents/skills/reality-verification in your project. Codex loads it when a task matches its description.
Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add tzachbon/smart-ralph --skill reality-verification -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/reality-verification, .gemini/skills/reality-verification, .github/skills/reality-verification and .opencode/skills/reality-verification in your project.
Going by SKILL.md and its folder, Reality Verification needs the command-line tools its instructions call (pnpm, gh and tsc).
SKILL.md contains no URLs. Its commands use gh, which can reach the network depending on how they are called. This is read from the text; nothing was executed.
Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.
Reality Verification is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.
About 877 tokens (SKILL.md is roughly 3.5k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full. Its references folder adds about 1.4k tokens, read only when the agent opens those files.
Skills that share tags, products or a category with Reality Verification: DeerFlow Smoke Test (bytedance/deer-flow, 84k stars), OpenWork Desktop CDP Driver (different-ai/openwork, 24k stars), E2E (gronxb/hot-updater, 1.8k stars) and Cloud Agents Starter (scalar/scalar, 16k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.
tzachbon (a GitHub user) maintains it in tzachbon/smart-ralph, which has 558 GitHub stars. The repository holds 26 skills in this directory. The repository was last updated on September 16, 2026.
Source: tzachbon/smart-ralph on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.