Agent skill

Omh Visual QA

by rlaope in rlaope/oh-my-hermes

[omh] Rendered UI needing a visual verdict: prepare observed-only rendered QA gates for web, frontend, image, document, and TUI surfaces.

MITAuto-check passedTesting & QA

Install Omh Visual QA

skills CLI
$ npx skills add rlaope/oh-my-hermes --skill omh-visual-qa -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install rlaope/oh-my-hermes omh-visual-qa --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/rlaope/oh-my-hermes.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/omh-visual-qa .claude/skills/omh-visual-qa && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
omh-visual-qa
GitHub stars
3.2k
Token cost
~3.2k tokens
SKILL.md length
1,510 words
Files
2 (incl. references)
Skills in repo
143
Repo updated
First seen
Licence
MIT

At a glance

[omh] Rendered UI needing a visual verdict: prepare observed-only rendered QA gates for web, frontend, image, document, and TUI surfaces.

  • The user says: visual-qa
  • SKILL.md covers Why This Exists, Do Not Use When, Examples and Completion Checklist, plus 5 more sections
  • Instructions only: no scripts, shell commands, URLs or credentials in SKILL.md
  • Visual quality assurance

What it does

Omh Visual QA is an agent skill from rlaope/oh-my-hermes. [omh] Rendered UI needing a visual verdict: prepare observed-only rendered QA gates for web, frontend, image, document, and TUI surfaces. Use when the user says: visual-qa, visual qa, visual QA, visual quality assurance, visual check, web qa, web visual qa, screenshot qa.

Its SKILL.md is about 3.2k tokens, which your agent loads only when the skill is triggered. The skill folder holds 2 other files, including reference files (for example `references/visual-verdict-contract.md`).

It sits in Testing & QA, covering Visual regression testing. The repository describes itself as: All in one plugin for Hermes Agent ⚚ the coding intelligence, a long-term memory system and model optimized workflow packages. The licence is MIT.

When your agent uses it

  • The user says: visual-qa
  • Visual quality assurance

Example prompts

  • “/omh-visual-qa”

What it can do on your machine

Read from SKILL.md and the folder at commit 41de9dc. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    No scripts in the folder and no shell commands in SKILL.md (its code samples are bash).

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Omh Visual QA loads about 3.2k tokens when it runs, and up to ~4.4k if it reads all its reference files. Until then it costs about 72 tokens; SKILL.md has 1,510 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~72
When it runs · the whole SKILL.md, loaded when a task matches
~3.2k
With references · SKILL.md plus every file in references/, read only if the agent opens them
~4.4k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from rlaope/oh-my-hermes at commit 41de9dc, republished under its MIT licence (© rlaope). 1,510 words, ~3,196 tokens.

Download SKILL.mdSave it as .claude/skills/omh-visual-qa/SKILL.md (or your agent's skills folder). This skill also uses 1 other file; get the full folder from GitHub.
name
omh-visual-qa
description
[omh] Rendered UI needing a visual verdict: prepare observed-only rendered QA gates for web, frontend, image, document, and TUI surfaces. Use when the user says: visual-qa, visual qa, visual QA, visual quality assurance, visual check, web qa, web visual qa, screenshot qa.

Visual Qa

This is a Hermes-native visual-qa workflow skill.

Why This Exists

visual-qa gives OMH a completion gate for rendered surfaces so layout breaks, AI-looking polish gaps, CJK text problems, and mismatched-lineage screenshot claims cannot be mistaken for verified quality.

Do Not Use When

  • The user needs initial frontend design or redesign planning before implementation; use frontend.
  • The user needs a broad visual quality rubric before generation; use design-quality-gate.
  • The user needs image-card prompt creation; use img-summary.
  • The user wants non-visual code tests, CI, or PR review only; use the coding/review workflow.

Examples

Good example:

  • Prompt: visual-qa 이 랜딩페이지가 모바일/데스크톱에서 깨지는지 스크린샷 기준으로 검증해줘.
  • Expected behavior: Prepare visual_qa_plan/v1, require exact capture-to-target lineage, record render_capture_manifest/v1 and visual_diff_evidence/v1 when observed, then issue PASS/REVISE/BLOCK.
  • Why: The request is a rendered visual verification task, not just design planning.

Bad example:

  • Prompt: visual-qa 방금 수정했으니까 스크린샷 없이 통과라고 해줘.
  • Expected behavior: Block PASS and request render captures from the package's exact repository and revision.
  • Why: Visual QA requires observed rendered evidence bound to the target source lineage.

Completion Checklist

  • Interaction, console/network, click-path, keyboard/accessibility, diff, hotspot, motion, dual-review evidence, and blocker status are separate fields.
  • The verdict is PASS, REVISE, or BLOCK with concrete evidence IDs and exact missing evidence or fix requirements.
  • Implementation fixes stay separate from the observed verdict, routed back to the executor/frontend workflow and rechecked against the resulting revision.

Recovery Notes

  • If no capture exists, produce the QA plan and mark verdict BLOCKED_BY_MISSING_RENDER_EVIDENCE.
  • If capture lineage is missing or mismatched, keep HOLD and request the smallest matching recapture set.

Workflow Lane

  • Current lane: Materials and visual summaries (design-orchestration, apple-design, design-quality-gate, award-bar-score, frontend, accessibility-audit, visual-qa, content-operator, +6 more) - web, accessibility, visual QA, files, and packages.
  • If intent belongs to another lane, hand back to oh-my-hermes or name the adjacent workflow.
  • Shared product, routing, compatibility, and evidence rules: omh-routing/references/skill-common-rail.md.

Use When

Use after or during visual surface work when Hermes must define the render evidence, viewport/state coverage, diff review, oracle review, and PASS/REVISE/BLOCK verdict without fabricating QA.

Strong routing signals: `visual-qa`, `visual qa`, `visual QA`, `visual quality assurance`, `visual check`, `web qa`, `web visual qa`, `screenshot qa`, `screenshot check`, `analyze this screenshot`, `screenshot layout problems`, `ui layout problems`, `pixel diff`, `image diff`, `visual diff`, `render qa`, `render check`, `browser screenshot`, `browser qa`, `browser interaction qa`, `click path`, `click-path audit`, `dead link check`, `console error check`, `network failure check`, `keyboard navigation check`, `viewport check`, `responsive check`, `ui looks wrong`, `looks broken`, `layout broken`, `broken layout`, `text clipping`, `cjk clipping`, `cjk layout`, `tui check`, `terminal ui check`, `スクリーンショットで確認`, `レイアウト崩れ`, `画面崩れ`, `見た目のQA`, `비주얼 qa`, `비주얼QA`, `시각 qa`, `시각 검증`, `화면 검증`, `스크린샷 검증`, `스크린샷 ui 레이아웃`, `스크린샷 UI 레이아웃`, `스크린샷 레이아웃 문제`, `렌더 검증`, `픽셀 diff`, `픽셀 비교`, `화면 깨짐`, `레이아웃 깨짐`, `글자 잘림`, `한글 줄바꿈`, `터미널 ui`, `截图检查`, `页面错位`, `视觉验收`, `布局错乱`

Catalog Metadata

Category: materials Phase: visual-qa Hermes role: operator Quality tier: visual-qa-gated Reasoning demand: standard

Quality bar:

  • List the exact pages, states, viewports, files, images, or TUI frames being checked.
  • For TUI surfaces, bind every capture to an explicit terminal size (80x24 and 120x40 at minimum); pasted rendered output at a named size is the screenshot-equivalent, and a capture without its size is not evidence.
  • Combine objective capture/diff evidence, hotspot review, alpha/transparent-background checks, and human-readable visual findings.
  • Capture interaction, click-path, and motion states when the UI has transitions or controls that change state.
  • Separate design-system consistency, functional integrity, visual fidelity, responsive behavior, accessibility visibility, and CJK/text precision.
  • Score every round through references/visual-verdict-contract.md: integer 0-100 score, PASS/REVISE/BLOCK, and a differences list pairing each observed problem with the smallest fix.
  • Hold 90 as the pass line: under it the verdict is REVISE and the named edits, a recapture of the same pages/states/viewports, and a fresh scored round are owed; rescoring the same captures is not a new round.
  • A host-collected sub-90 baseline needs a changed revision, next round ordinal, and newer same-condition capture; plan caps only tighten.

Handoff policy:

Keep the QA plan, evidence manifest, target-lineage rule, and verdict narration in Hermes. Screenshots, TUI captures, image diffs, browser runs, OCR/CJK checks, and oracle reviews are observed evidence supplied by the wrapper, executor, or user.

Required inputs:

  • surface type
  • target URL, route, file, image, or TUI command when available
  • intended design, baseline, or reference
  • pages, states, viewports, and locales to cover
  • complete page/state/viewport enumeration rather than a sample
  • target repository and exact source revision
  • known risk areas such as CJK, overflow, responsiveness, or accessibility
  • motion and interaction states that need capture
  • browser interaction paths, mutating-flow boundary, and test credentials policy when a live web UI is in scope
  • console, network, accessibility, and keyboard navigation checks required for browser QA claims
  • render/capture evidence bound to the target repository and revision for completion claims

Expected outputs:

  • visual_qa_plan/v1
  • web_visual_qa_package/v2
  • viewport_state_capture_matrix/v1
  • message_attachment_projection/v1 for chat attachments
  • web_visual_qa_message_card/v1 for chat message summaries
  • render_capture_manifest/v1 when observed
  • browser_interaction_trace/v1 when observed
  • console_network_health/v1 when observed
  • click_path_state_trace/v1 when observed
  • accessibility_keyboard_trace/v1 when observed
  • visual_diff_evidence/v1 when observed
  • visual_hotspot_review/v1 when observed
  • motion_interaction_capture/v1 when observed
  • dual_oracle_visual_review/v1 when observed
  • cjk_layout_findings/v1 when applicable
  • visual_qa_verdict/v1
  • retry_or_blocker/v1

Artifact expectations:

  • visual_qa_plan/v1 with pages, states, viewports, references, and exact target repository/revision lineage
  • web_visual_qa_package/v2 with target_lineage, unique required_viewports, capture source_lineage, blocking_violations, criteria, reviews, auto routing, and observed-only cost policy
  • viewport_state_capture_matrix/v1 enumerates every route/page, 375/768/1280-style viewport, scroll position, modal/tab state, and CJK-heavy region to capture
  • message_attachment_projection/v1 maps eligible observed captures to attachment candidates without claiming delivery
  • web_visual_qa_message_card/v1 projects recorded criteria, captures, routing, cost policy, and attachment hints into chat-safe copy
  • render_capture_manifest/v1 only from captures whose source lineage matches the target package
  • browser_interaction_trace/v1 only from observed journey runs with read-only or staging-safe boundaries recorded
  • console_network_health/v1 records observed console errors, failed requests, status codes, and ignored third-party noise
  • click_path_state_trace/v1 maps each touchpoint to its handler, state reads/writes, final UI state, and undo/race/stale-closure risks
  • accessibility_keyboard_trace/v1 records observed focus order, keyboard reachability, and automated scan boundaries
  • visual_diff_evidence/v1 only when the wrapper/executor records objective diff output such as dimensionsMatch, diffRatio, similarityScore, alphaChannelIntact, and hotspots
  • motion_interaction_capture/v1 only when motion frames are observed before, during, and after transition
  • visual_hotspot_review/v1 maps diff hotspots, TUI overflow lines, or screenshot regions to visual causes
  • dual_oracle_visual_review/v1 only when independent read-only review evidence exists
  • visual_qa_verdict/v1 with the integer 0-100 score, PASS/REVISE/BLOCK, and difference/suggestion pairs
  • PASS unavailable until capture repository/revision lineage exactly matches the package target, every required viewport is captured, and all supplied blocking findings are resolved
  • web_qa_observation_run/v1 and web_qa_comparison/v1 only from a host_web_qa_adapter_receipt/v1 imported through omh web-qa observation: seven independently observed channels or a named blocker per cell
Show full SKILL.md (444 more words)Show less

Safety rules:

  • Never claim PASS without rendered evidence whose repository and revision exactly match the package target lineage.
  • Source review, mismatched-lineage captures, generated plans, and unobserved browser commands are not visual QA evidence.
  • Do not sample only one good page, viewport, or state when the surface has more; missed pages, modals, scroll states, or CJK-heavy regions keep PASS unavailable.
  • Do not run destructive browser journeys such as checkout, payment, delete, or mass-update on production URLs; require staging or explicit safe test boundaries and redact credentials/PII from captures.
  • Do not claim browser interaction PASS without observed click-path/state-transition traces for the touchpoints in scope.
  • Do not claim accessibility from automated scan output alone; keyboard and focus-order evidence are separate observed checks.
  • Pixel diff localizes hotspots only; it never produces the score or verdict, and objective diffs are evidence, not verdicts: review visual hierarchy, layout, CJK text, state coverage, and product intent separately.
  • Do not excuse diff hotspots as animation; capture settled frames and motion frames separately.
  • Claim high confidence only with two read-only reviews: design-system/functional integrity and visual fidelity/CJK precision.
  • Operator-supplied blocking criteria (CJK clipping, broken wrapping, overlapping UI, invisible text, unusable controls, offscreen critical content) block PASS until _validate_pass sees passing evidence refs.
  • Do not launch, poll, or watch browsers, image tools, LLMs, or external services from OMH core; the selected host or executor adapter does that work.
  • A host receipt is observation, not permission: a missing channel keeps BLOCK, unequal condition digests are not_comparable, and a completed run is reused, not recollected.

Runtime Evidence

Preferred harness for this skill: visual-qa.

sh
omh runtime record --skill visual-qa --harness visual-qa --status started

Record observed delegation results; otherwise return not_available or not_observed. Prepared OMH routing is not execution, review, CI, merge-readiness, or merge evidence.

  • Treat wrapper memory/context summaries as advisory local context, not proof of opaque Hermes memory reads or changes. Preserve workflow intent and stop conditions; verify before claiming completion. Reply in the user's own words and the host's own voice: its SOUL.md persona owns reply language, tone, speech level, and sentence endings, progress updates included (where it sets no language, use the one the user wrote in), and OMH shapes structure and content only; OMH's record terms (surface, lane, wrapper, handoff, evidence boundary, not_observed) stay in records and tool calls, never in the sentence the user reads unless they ask about one; and when a stop condition or a decision the user owns ends the turn, offer the next action as a question rather than declaring what will not be done.

Use Hermes-native subagent/delegation features when available: native subagents -> Hermes delegation when available, otherwise sequential lanes.

Shared product, compatibility, topology, memory, harness, and execution rules: omh-routing/references/skill-common-rail.md. Load it when applicable; otherwise name an unavailable capability.

© rlaope, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

SKILL.md and 1 other file (references) in skills/omh-visual-qa of rlaope/oh-my-hermes.

  • SKILL.md
  • references/visual-verdict-contract.md

Open the folder on GitHubat commit 41de9dc

Compare with similar skills

Omh Visual QA next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Omh Visual QA compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Omh Visual QA this skillrlaope/oh-my-hermes3.2k—~3.2kAutomated safety check: PassMIT
Review Local UI Screenshotsxiaocang/easydict_win32103—~1.1kAutomated safety check: PassGPL-3.0
Handsontable Visual Test Demoshandsontable/handsontable22k—~1.3kAutomated safety check: PassCustom licence
Kc Screenshotimran31415/kube-coder388—~870Automated safety check: NotesMIT
Visual QAliangdabiao/Godogen126—~1.4kAutomated safety check: PassMIT
Meticulous FixFlintSH/Flare135—~2.3kAutomated safety check: PassMIT

Similar skills

  • Review Local UI Screenshots

    xiaocang/easydict_win32

    Review Easydict UI automation screenshot artifacts already present in local artifacts/ui-screenshots, screenshots, or a user-provided artifact directory.

    103 GitHub stars~1.1k tokensUpdated today
    Testing & QAAuto-check passed
  • Handsontable Visual Test Demos

    handsontable/handsontable

    Explains how to add or change the demo pages that Handsontable's visual regression suite photographs, including per-feature routes in the js demo and the shared grid.

    22k GitHub stars~1.3k tokensUpdated today
    Testing & QAAuto-check passed
  • Kc Screenshot

    imran31415/kube-coder

    Capture desktop + mobile, dark + light screenshots of the kube-coder dashboard SPA for visual QA of a UI change.

    388 GitHub stars~870 tokensUpdated today
    Testing & QAAuto-check: notes
  • Visual QA

    liangdabiao/Godogen

    Visual quality assurance — analyze game screenshots for defects, compare against reference, check motion in frame sequences.

    126 GitHub stars~1.4k tokensUpdated 5 mo ago
    Testing & QAAuto-check passed
  • Meticulous Fix

    FlintSH/Flare

    Fix the visual diffs that have been reviewed and rejected on a Meticulous test run, following their review comments if given.

    135 GitHub stars~2.3k tokensUpdated 3 days ago
    Testing & QAAuto-check passed
  • Visual QA For Web And Terminal UIs

    code-yeongyu/oh-my-openagent

    Checks a built UI against a human-interface checklist and any reference image, pairing day and night screenshots at two screen widths.

    70k GitHub stars~9.5k tokensUpdated today
    Testing & QAAuto-check passed

More from rlaope/oh-my-hermes

All 143 skills in this repo
  • Omh Accessibility Audit

    rlaope/oh-my-hermes

    [omh] Screen-reader or keyboard accessibility gaps: prepare WCAG, keyboard, focus, screen-reader, target-size, and reflow evidence gates for UI surfaces.

    3.2k GitHub stars~2.8k tokensUpdated today
    Auto-check passed
  • Omh Agent Evaluation

    rlaope/oh-my-hermes

    [omh] Choosing between coding agents on evidence: compare executor or agent choices on reproducible tasks using quality, cost, time, tool, and evidence metrics.

    3.2k GitHub stars~2.1k tokensUpdated today
    Auto-check passed
  • Omh Agent Instructions

    rlaope/oh-my-hermes

    [omh] Agent instruction file for a repo -- AGENTS.md, CLAUDE.md, a Cursor rule: write or update what an agent cannot derive from the code, inside a marked region, with every command verified or…

    3.2k GitHub stars~2.2k tokensUpdated today
    Auto-check passed
  • Omh Agent Ops Review

    rlaope/oh-my-hermes

    [omh] AI agent progress for managers: help managers inspect AI-agent progress, blockers, quality gates, and throughput levers.

    3.2k GitHub stars~1.9k tokensUpdated today
    Auto-check passed
  • Omh AI Slop Cleaner

    rlaope/oh-my-hermes

    [omh] Messy or AI-generated code to clean up: delete AI-generated slop, dead code, and duplication while observable behavior stays identical.

    3.2k GitHub stars~2.7k tokensUpdated today
    Auto-check passed
  • Omh App Debugging

    rlaope/oh-my-hermes

    [omh] Application code misbehaves -- a wrong value, a flaky test, a lost update: reproduce it first, form competing hypotheses, discriminate them with the cheapest observation, and only then fix the…

    3.2k GitHub stars~2.3k tokensUpdated today
    Auto-check passed

Categories

Questions about Omh Visual QA

What does Omh Visual QA do?

[omh] Rendered UI needing a visual verdict: prepare observed-only rendered QA gates for web, frontend, image, document, and TUI surfaces. Omh Visual QA is an agent skill from rlaope/oh-my-hermes. [omh] Rendered UI needing a visual verdict: prepare observed-only rendered QA gates for web, frontend, image, document, and TUI surfaces.

When should I use Omh Visual QA?

Omh Visual QA fits situations like: the user says: visual-qa; visual quality assurance.

How do I install Omh Visual QA in Claude Code?

Run `npx skills add rlaope/oh-my-hermes --skill omh-visual-qa -a claude-code`. Or copy the skill folder (skills/omh-visual-qa in rlaope/oh-my-hermes) into .claude/skills/omh-visual-qa in your project. Claude Code loads it when a task matches its description.

How do I install Omh Visual QA in Codex?

Run `npx skills add rlaope/oh-my-hermes --skill omh-visual-qa -a codex`. Or copy the skill folder (skills/omh-visual-qa in rlaope/oh-my-hermes) into .agents/skills/omh-visual-qa in your project. Codex loads it when a task matches its description.

Can I use Omh Visual QA in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add rlaope/oh-my-hermes --skill omh-visual-qa -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/omh-visual-qa, .gemini/skills/omh-visual-qa, .github/skills/omh-visual-qa and .opencode/skills/omh-visual-qa in your project.

What does Omh Visual QA need to run?

SKILL.md names no scripts, command-line tools or credentials: Omh Visual QA is instructions for the agent only.

Does Omh Visual QA access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is Omh Visual QA safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Omh Visual QA use?

Omh Visual QA is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Omh Visual QA use?

About 3.2k tokens (SKILL.md is roughly 13k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full. Its references folder adds about 1.2k tokens, read only when the agent opens those files.

What are the alternatives to Omh Visual QA?

Skills that share tags, products or a category with Omh Visual QA: Review Local UI Screenshots (xiaocang/easydict_win32, 103 stars), Handsontable Visual Test Demos (handsontable/handsontable, 22k stars), Kc Screenshot (imran31415/kube-coder, 388 stars) and Visual QA (liangdabiao/Godogen, 126 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Omh Visual QA?

rlaope (a GitHub user) maintains it in rlaope/oh-my-hermes, which has 3,233 GitHub stars. The repository holds 143 skills in this directory. The repository was last updated on October 8, 2026.

Source: rlaope/oh-my-hermes on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.