Diff-Driven Smoke Tests
Skyvern-AI/skyvern
Reads your git diff, writes a handful of happy-path browser smoke tests, runs them with Skyvern or Chrome DevTools MCP and posts screenshot evidence to the PR.
Sweeps a running web app by clicking every reachable control, then reports dead buttons, console errors, failed requests and mismatches between API data and the screen.
$ npx skills add reticlehq/reticle --skill audit-my-app -a claude-codeProject install by default; add -g for ~/.claude/skills/.
$ gh skill install reticlehq/reticle audit-my-app --agent claude-codeProject scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).
$ git clone --depth 1 https://github.com/reticlehq/reticle.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/audit-my-app .claude/skills/audit-my-app && rm -rf skills-srcUse ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.
Claude Code skills documentation · loads skills from .claude/skills/
Install the "audit-my-app" agent skill from https://github.com/reticlehq/reticle/tree/main/skills/audit-my-app into .claude/skills/audit-my-app/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "audit-my-app", then confirm the skill loads.Claude Code copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$skill-installer install https://github.com/reticlehq/reticle/tree/main/skills/audit-my-appType this inside Codex. $skill-installer <name> installs a curated skill from openai/skills. The installer writes to $CODEX_HOME/skills (default ~/.codex/skills). Restart Codex if the skill does not show up.
$ npx skills add reticlehq/reticle --skill audit-my-app -a codexProject install goes to .agents/skills/; add -g for ~/.codex/skills/.
$ gh skill install reticlehq/reticle audit-my-app --agent codexProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/reticlehq/reticle.git skills-src && mkdir -p .agents/skills && cp -r skills-src/skills/audit-my-app .agents/skills/audit-my-app && rm -rf skills-srcUse ~/.agents/skills/ instead of .agents/skills for a personal install.
Codex skills documentation · loads skills from .agents/skills/
Install the "audit-my-app" agent skill from https://github.com/reticlehq/reticle/tree/main/skills/audit-my-app into .agents/skills/audit-my-app/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "audit-my-app", then confirm the skill loads.Codex copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add reticlehq/reticle --skill audit-my-app -a cursorProject install goes to .agents/skills/; add -g for ~/.cursor/skills/.
$ gh skill install reticlehq/reticle audit-my-app --agent cursorProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/reticlehq/reticle.git skills-src && mkdir -p .cursor/skills && cp -r skills-src/skills/audit-my-app .cursor/skills/audit-my-app && rm -rf skills-srcUse ~/.cursor/skills/ instead of .cursor/skills for a personal install.
Cursor skills documentation · loads skills from .cursor/skills/, .agents/skills/, .claude/skills/, .codex/skills/
Install the "audit-my-app" agent skill from https://github.com/reticlehq/reticle/tree/main/skills/audit-my-app into .cursor/skills/audit-my-app/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "audit-my-app", then confirm the skill loads.Cursor copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gemini skills install https://github.com/reticlehq/reticle.git --path skills/audit-my-app--scope user (default) or --scope workspace; --path is the subfolder of the repo that holds the skill; --consent skips the security confirmation prompt.
$ npx skills add reticlehq/reticle --skill audit-my-app -a gemini-cliProject install goes to .agents/skills/; add -g for ~/.gemini/skills/.
$ gh skill install reticlehq/reticle audit-my-app --agent gemini-cliProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/reticlehq/reticle.git skills-src && mkdir -p .gemini/skills && cp -r skills-src/skills/audit-my-app .gemini/skills/audit-my-app && rm -rf skills-srcUse ~/.gemini/skills/ instead of .gemini/skills for a personal install, then run /skills reload.
Gemini CLI skills documentation · loads skills from .gemini/skills/, .agents/skills/
Install the "audit-my-app" agent skill from https://github.com/reticlehq/reticle/tree/main/skills/audit-my-app into .gemini/skills/audit-my-app/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "audit-my-app", then confirm the skill loads.Gemini CLI copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gh skill install reticlehq/reticle audit-my-appInstalls for Copilot at project scope by default; add --scope user for a personal install. Preview a skill first with gh skill preview. Needs GitHub CLI 2.90.0 or later (public preview).
$ npx skills add reticlehq/reticle --skill audit-my-app -a github-copilotProject install goes to .agents/skills/; add -g for ~/.copilot/skills/.
$ git clone --depth 1 https://github.com/reticlehq/reticle.git skills-src && mkdir -p .github/skills && cp -r skills-src/skills/audit-my-app .github/skills/audit-my-app && rm -rf skills-srcUse ~/.copilot/skills/ instead of .github/skills for a personal install. Commit .github/skills so cloud agent and code review can use it.
GitHub Copilot skills documentation · loads skills from .github/skills/, .claude/skills/, .agents/skills/
Install the "audit-my-app" agent skill from https://github.com/reticlehq/reticle/tree/main/skills/audit-my-app into .github/skills/audit-my-app/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "audit-my-app", then confirm the skill loads.GitHub Copilot copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add reticlehq/reticle --skill audit-my-app -a opencodeOpenCode documents no install command of its own. Project install goes to .agents/skills/; add -g for ~/.config/opencode/skills/.
$ gh skill install reticlehq/reticle audit-my-app --agent opencodeProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/reticlehq/reticle.git skills-src && mkdir -p .opencode/skills && cp -r skills-src/skills/audit-my-app .opencode/skills/audit-my-app && rm -rf skills-srcUse ~/.config/opencode/skills/ instead of .opencode/skills for a personal install.
OpenCode skills documentation · loads skills from .opencode/skills/, .claude/skills/, .agents/skills/
Install the "audit-my-app" agent skill from https://github.com/reticlehq/reticle/tree/main/skills/audit-my-app into .opencode/skills/audit-my-app/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "audit-my-app", then confirm the skill loads.OpenCode copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
audit-my-appSweeps a running web app by clicking every reachable control, then reports dead buttons, console errors, failed requests and mismatches between API data and the screen.
The skill uses Reticle to check an app without needing to understand its code. It first asks the app to describe its own testable surface (test ids, domain signals, stores and saved flows), then runs a crawl that clicks everything reachable, bounded by a step limit that defaults to 25. Two counts signal real problems: dead controls, meaning buttons wired to nothing, and contradictions, where a channel disagrees with what the screen showed. Console errors and failed requests are worth reading but a busy app produces them innocently.
Because the crawl clicks everything, it should point at a development environment, and a separate explore call gives a non-destructive pass over what is reachable. A reconcile step compares what the API returned with what rendered, such as ten rows from the API but nine in the table, which neither the network log nor the DOM shows alone. Coverage and domain calls list the parts nobody exercised and whether saved flows actually assert anything. The skill warns against hand-rolling the sweep, since a dead button throws no error.
5 steps, taken from the step headings in SKILL.md.
Read from SKILL.md and the folder at commit 178e5c0. It shows what the files ask for, not the result of running them.
Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.
From allowed-tools in the SKILL.md frontmatter.
Shell commands in SKILL.md call:
curlnpxFrom the folder's file list and the shell code blocks in SKILL.md.
Hosts in commands or code, which the agent is likely to contact:
docs.reticle.shFrom URLs in SKILL.md, links to its own repository left out.
Names no API keys, tokens, secrets or passwords.
From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.
Whole-App Health Sweep loads about 1.1k tokens when it runs. Until then it costs about 113 tokens; SKILL.md has 423 words of instructions outside code blocks.
Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.
The automated check found no risky patterns in SKILL.md.
Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.
The full file from reticlehq/reticle at commit 178e5c0, republished under its Apache-2.0 licence (© reticlehq). 423 words, ~1,072 tokens.
.claude/skills/audit-my-app/SKILL.md (or your agent's skills folder).You do not need to understand the codebase to check it. Reticle drives every reachable control in the running app and reports the anomalies. Not installed? RETICLE_INSTALL_SOURCE=npx_skill npx @reticlehq/server@latest init, then the install-and-verify skill.
reticle_run({ tool: "reticle_capabilities", args: { sessionId } })About 1 KB, and it is the app describing its own testable surface: every registered testid, every domain signal, the stores, and the saved flows with their steps. That beats snapshotting the DOM and inferring intent from element names, and it is the cheapest orientation available.
reticle_run({ tool: "reticle_verify", sessionId, args: { action: "crawl", maxSteps: 25 } }){
"interactiveFound": 3,
"stepsRun": 3,
"anomalies": [],
"counts": { "consoleErrors": 0, "failedRequests": 0, "deadControls": 0, "contradictions": 0 },
"visited": ["- textbox \"Email\"", "- button \"Sign in\""],
"truncated": false
}deadControls and contradictions are the two counts that mean a real problem. Console errors and failed requests are worth reading but a busy app produces both innocently. A dead control is a button wired to nothing; a contradiction is a channel disagreeing with what the screen showed.
It clicks everything, so point it at a dev environment. maxSteps bounds it and defaults to 25. Want a non-destructive pass first: what is reachable, without touching it? reticle_run({ tool: "reticle_explore", sessionId }).
Do not hand-roll this sweep. The obvious version (click each control, assert no console error) passes on exactly the bug you are sweeping for, because a dead button throws nothing.
reticle_run({ tool: "reticle_reconcile", sessionId })The API returned ten rows, the table shows nine, nothing errored. Neither the network log nor the DOM is wrong on its own. Only the comparison catches it, and nothing else you can run makes that comparison.
reticle_run({ tool: "reticle_verify", sessionId, args: { action: "coverage" } }) // { total, exercised, untouched }And if the project already has saved flows, ask whether they prove anything:
reticle_run({ tool: "reticle_domain", sessionId })
// → { flowCount, coverage: { asserted, presenceOnly, assertionFree }, gaps: { declaredUntestedSignals, … } }A suite of forty flows where thirty-one assert nothing is a suite that will stay green through any regression. That number is usually the most alarming thing in the whole audit, and nothing else reports it.
Lead with the counts, then one line per real finding with its file:line from reticle_look { action: "element" }. Separate:
untouched controls and assertionFree flows. Not known to be broken; known to be unchecked.Do not report a clean audit over a partial one. If truncated is true or maxSteps cut the sweep short, say what was not reached. A silent cap reads as "everything is fine" when it means "I stopped".
Full capability reference: curl https://docs.reticle.sh/capabilities.md. Everything else: curl https://docs.reticle.sh/llms.txt.
© reticlehq, Apache-2.0. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file
Just SKILL.md in skills/audit-my-app of reticlehq/reticle.
Open the folder on GitHubat commit 178e5c0
Whole-App Health Sweep next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.
| Skill | Stars | Used in | Tokens | Auto-check | Licence | Repo updated |
|---|---|---|---|---|---|---|
| Whole-App Health Sweep this skillreticlehq/reticle | 1.2k | — | ~1.1k | Automated safety check: Pass | Apache-2.0 | |
| Diff-Driven Smoke TestsSkyvern-AI/skyvern | 23k | — | ~5.2k | Automated safety check: Pass | AGPL-3.0 | |
| LangBot Testinglangbot-app/LangBot | 18k | — | ~1k | Automated safety check: Notes | Apache-2.0 | |
| Agentic Browser Testingpetrkindlmann/qa-skills | 170 | — | ~4.5k | Automated safety check: Pass | MIT | |
| Testing QAaiskillstore/marketplace | 433 | 3 repos | ~1.2k | Automated safety check: Pass | None | |
| Playwright E2E Testsonyx-dot-app/onyx | 32k | 1 repos | ~2.8k | Automated safety check: Notes | Custom licence |
Skyvern-AI/skyvern
Reads your git diff, writes a handful of happy-path browser smoke tests, runs them with Skyvern or Chrome DevTools MCP and posts screenshot evidence to the PR.
langbot-app/LangBot
Tests LangBot's WebUI and core flows through an automated browser and backend logs, with a routing table to reference guides per feature area.
petrkindlmann/qa-skills
Goal-driven E2E testing where a browser agent (Playwright MCP / computer-use) reads a natural-language goal and explores the app via the accessibility tree to assert outcomes — no pre-written script.
aiskillstore/marketplace
Comprehensive testing and QA workflow covering unit testing, integration testing, E2E testing, browser automation, and quality assurance.
onyx-dot-app/onyx
Write and maintain Playwright end-to-end tests for the Onyx application.
ktnyt/cclsp
Performs manual hands-on testing of a web application using playwright-cli.
reticlehq/reticle
Applies red-green TDD to behavior unit tests cannot reach, by stating the expected outcome against the running app with Reticle before writing the feature.
reticlehq/reticle
Finds why a running web app misbehaves when the console is empty and the code looks fine, by reading the click, request, store and console together.
reticlehq/reticle
Drives and verifies Electron or Tauri desktop apps through Reticle, which sees the renderer and the IPC calls that a browser-based testing tool cannot observe.
reticlehq/reticle
Finds out why a passing test suite sits on top of a broken app by comparing what the running app does with what the tests claim, using Reticle.
reticlehq/reticle
Picks up bugs a person flagged by pointing at elements in the running app, each mark carrying the element, their note and the source file and line, then fixes and verifies them.
reticlehq/reticle
Saves a user journey driven through the app as a deterministic regression check that replays with no model and no test code, using Reticle.
Works with
Categories
Sweeps a running web app by clicking every reachable control, then reports dead buttons, console errors, failed requests and mismatches between API data and the screen. The skill uses Reticle to check an app without needing to understand its code. It first asks the app to describe its own testable surface (test ids, domain signals, stores and saved flows), then runs a crawl that clicks everything reachable, bounded by a step limit that defaults to 25.
Whole-App Health Sweep fits situations like: smoke testing an unfamiliar codebase without writing a script; checking an app after a large merge or dependency bump; running a pre-release health check on a web app; finding screens whose displayed data does not match the API response.
Run `npx skills add reticlehq/reticle --skill audit-my-app -a claude-code`. Or copy the skill folder (skills/audit-my-app in reticlehq/reticle) into .claude/skills/audit-my-app in your project. Claude Code loads it when a task matches its description.
Run `npx skills add reticlehq/reticle --skill audit-my-app -a codex`. Or copy the skill folder (skills/audit-my-app in reticlehq/reticle) into .agents/skills/audit-my-app in your project. Codex loads it when a task matches its description.
Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add reticlehq/reticle --skill audit-my-app -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/audit-my-app, .gemini/skills/audit-my-app, .github/skills/audit-my-app and .opencode/skills/audit-my-app in your project.
Going by SKILL.md and its folder, Whole-App Health Sweep needs the command-line tools its instructions call (curl and npx). Our summary lists: A running web app in a development environment; Reticle installed in the project through npx.
SKILL.md names 1 domain. In commands or code: docs.reticle.sh; the agent is likely to contact it when it follows the instructions. This is read from the text; nothing was executed.
Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.
Whole-App Health Sweep is published under the Apache-2.0 licence (declared in SKILL.md). It allows redistribution, so the full SKILL.md is shown on this page.
About 1.1k tokens (SKILL.md is roughly 4.3k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.
Skills that share tags, products or a category with Whole-App Health Sweep: Diff-Driven Smoke Tests (Skyvern-AI/skyvern, 23k stars), LangBot Testing (langbot-app/LangBot, 18k stars), Agentic Browser Testing (petrkindlmann/qa-skills, 170 stars) and Testing QA (aiskillstore/marketplace, 433 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.
reticlehq (a GitHub organization) maintains it in reticlehq/reticle, which has 1,199 GitHub stars. The repository holds 19 skills in this directory. The repository was last updated on October 9, 2026.
Source: reticlehq/reticle on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.