Agent skill

Visual Regression

by WrongStack in WrongStack/WrongStack

Detect unintended UI changes with reproducible rendered screenshots and reviewed baselines.

MITAuto-check passedTesting & QA

Install Visual Regression

skills CLI
$ npx skills add WrongStack/WrongStack --skill visual-regression -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install WrongStack/WrongStack visual-regression --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/WrongStack/WrongStack.git skills-src && mkdir -p .claude/skills && cp -r skills-src/packages/core/skills/visual-regression .claude/skills/visual-regression && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
visual-regression
GitHub stars
371
Token cost
~750 tokens
SKILL.md length
269 words
Files
1
Skills in repo
100
Repo updated
First seen
Licence
MIT

At a glance

Detect unintended UI changes with reproducible rendered screenshots and reviewed baselines.

  • Works in 6 steps: Pin viewport, browser, device scale,… → Wait for actual readiness:… → Mask only intentionally variable… → …
  • Adding screenshot checks
  • SKILL.md covers Selection card, Overview, Rules and Workflow, plus 3 more sections
  • Instructions only: no scripts, shell commands, URLs or credentials in SKILL.md

What it does

Visual Regression is an agent skill from WrongStack/WrongStack. Detect unintended UI changes with reproducible rendered screenshots and reviewed baselines. Use when adding screenshot checks or investigating layout/theme regressions; do not approve a changed image merely to make the test green.

Its SKILL.md is about 750 tokens, which your agent loads only when the skill is triggered. It is a single SKILL.md file with no bundled scripts.

It sits in Testing & QA, covering Visual regression testing. The repository describes itself as: An AI coding agent that reads your code, edits files, runs commands, and reasons through bugs — across a terminal REPL, a full-screen TUI, and a browser UI, while you keep your… The licence is MIT.

When your agent uses it

  • Adding screenshot checks
  • Investigating layout/theme regressions
  • Do not approve a changed image merely to make the test green

Example prompts

  • “/visual-regression”

Workflow steps

6 steps, taken from the first numbered list in SKILL.md.

  1. Pin viewport, browser, device scale, locale, theme, fonts, test data and rendered state.
  2. Wait for actual readiness: fonts/assets/data and known animation completion. Arbitrary sleeps make noisy baselines.
  3. Mask only intentionally variable content, with a documented reason; masking the changed feature defeats the test.
  4. Separate antialiasing/environment noise from a meaningful geometry, content or state change.
  5. Review each changed baseline against the requested outcome and prior image before accepting it.
  6. Keep screenshot assertions alongside behavioral checks; pixels alone do not prove accessible or functioning controls.

What it can do on your machine

Read from SKILL.md and the folder at commit a744bdc. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    No scripts in the folder and no shell commands in SKILL.md.

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    Links to these hosts (documentation or services it may open):

    • playwright.dev
    • storybook.js.org

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Visual Regression loads about 750 tokens when it runs. Until then it costs about 62 tokens; SKILL.md has 269 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~62
When it runs · the whole SKILL.md, loaded when a task matches
~750

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from WrongStack/WrongStack at commit a744bdc, republished under its MIT licence (© WrongStack). 269 words, ~750 tokens.

Download SKILL.mdSave it as .claude/skills/visual-regression/SKILL.md (or your agent's skills folder).
name
visual-regression
description
Detect unintended UI changes with reproducible rendered screenshots and reviewed baselines. Use when adding screenshot checks or investigating layout/theme regressions; do not approve a changed image merely to make the test green.
trigger
Detect unintended UI changes with reproducible rendered screenshots and reviewed baselines. Use when adding screenshot checks or investigating layout/theme…
version
1.0.1
required-capabilities
filesystem.read
optional-capabilities
filesystem.write, execution.shell, verification.run, web.research
metadata.routing-group
design
metadata.domain
design

Visual Regression

Selection card

  • Task: Compare rendered UI states against screenshot baselines. / TR: Render edilmiş arayüzü ekran görüntüsü baseline ile karşılaştır.
  • Start: Identify the target surface, reference, user task and existing tokens.
  • Finish: apply the acceptance checks below; report observed results and unresolved constraints.

Overview

Detect unintended UI changes with reproducible rendered screenshots and reviewed baselines.

Current checked targets: Playwright 1.64.0 and Storybook React 10.6.1 (2026-10-09). Refresh registry and browser/tool compatibility before installation.

Rules

  1. Pin viewport, browser, device scale, locale, theme, fonts, test data and rendered state.
  2. Wait for actual readiness: fonts/assets/data and known animation completion. Arbitrary sleeps make noisy baselines.
  3. Mask only intentionally variable content, with a documented reason; masking the changed feature defeats the test.
  4. Separate antialiasing/environment noise from a meaningful geometry, content or state change.
  5. Review each changed baseline against the requested outcome and prior image before accepting it.
  6. Keep screenshot assertions alongside behavioral checks; pixels alone do not prove accessible or functioning controls.

Workflow

  1. Choose representative routes/components and the states most likely to regress.
  2. Capture a deterministic baseline through the configured browser/story workflow.
  3. Inspect the diff image and trace the owning component/theme change.
  4. Fix the source or accept an intended baseline change with evidence.
  5. Rerun the same conditions, then the affected interaction tests.

Before returning

Inputs and baseline identity recorded; differences inspected; masking justified; intended changes distinguished from regressions.

Sources

Versioned facts checked 2026-10-09; refresh authoritative sources before new installs/upgrades. Playwright snapshots, Storybook visual testing.

Skills in scope

  • testing — behavioral and failure-path verification.
  • design-critique — rendered evidence and ranked findings.
  • accessibility — keyboard and assistive interaction.

© WrongStack, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

Just SKILL.md in packages/core/skills/visual-regression of WrongStack/WrongStack.

Open the folder on GitHubat commit a744bdc

Compare with similar skills

Visual Regression next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Visual Regression compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Visual Regression this skillWrongStack/WrongStack371—~750Automated safety check: PassMIT
Review Local UI Screenshotsxiaocang/easydict_win32104—~1.1kAutomated safety check: PassGPL-3.0
Handsontable Visual Test Demoshandsontable/handsontable22k—~1.3kAutomated safety check: PassCustom licence
Kc Screenshotimran31415/kube-coder388—~870Automated safety check: NotesMIT
Visual QAliangdabiao/Godogen126—~1.4kAutomated safety check: PassMIT
Meticulous FixFlintSH/Flare135—~2.3kAutomated safety check: PassMIT

Similar skills

  • Review Local UI Screenshots

    xiaocang/easydict_win32

    Review Easydict UI automation screenshot artifacts already present in local artifacts/ui-screenshots, screenshots, or a user-provided artifact directory.

    104 GitHub stars~1.1k tokensUpdated 2 days ago
    Testing & QAAuto-check passed
  • Handsontable Visual Test Demos

    handsontable/handsontable

    Explains how to add or change the demo pages that Handsontable's visual regression suite photographs, including per-feature routes in the js demo and the shared grid.

    22k GitHub stars~1.3k tokensUpdated yesterday
    Testing & QAAuto-check passed
  • Kc Screenshot

    imran31415/kube-coder

    Capture desktop + mobile, dark + light screenshots of the kube-coder dashboard SPA for visual QA of a UI change.

    388 GitHub stars~870 tokensUpdated 2 days ago
    Testing & QAAuto-check: notes
  • Visual QA

    liangdabiao/Godogen

    Visual quality assurance — analyze game screenshots for defects, compare against reference, check motion in frame sequences.

    126 GitHub stars~1.4k tokensUpdated 5 mo ago
    Testing & QAAuto-check passed
  • Meticulous Fix

    FlintSH/Flare

    Fix the visual diffs that have been reviewed and rejected on a Meticulous test run, following their review comments if given.

    135 GitHub stars~2.3k tokensUpdated 5 days ago
    Testing & QAAuto-check passed
  • Visual QA For Web And Terminal UIs

    code-yeongyu/oh-my-openagent

    Checks a built UI against a human-interface checklist and any reference image, pairing day and night screenshots at two screen widths.

    70k GitHub stars~9.5k tokensUpdated today
    Testing & QAAuto-check passed

More from WrongStack/WrongStack

All 100 skills in this repo
  • Tech Stack

    WrongStack/WrongStack

    Validate and upgrade dependencies against live registries and official migration guides in any ecosystem.

    374 GitHub stars~1.5k tokensUpdated today
    Auto-check passed
  • Skill Creator

    WrongStack/WrongStack

    Create, improve and validate WrongStack SKILL.md bundles with precise discovery, progressive resources and current runtime contracts.

    374 GitHub stars~1.3k tokensUpdated today
    Auto-check passed
  • Bug Hunter

    WrongStack/WrongStack

    A skill your agent uses when scanning source code for bugs, anti-patterns, code smells, or quality issues in a codebase, or when running a proof-driven bug hunt that must find, prove, fix, and…

    374 GitHub stars~2.4k tokensUpdated today
    Auto-check passed
  • Design Craft

    WrongStack/WrongStack

    Design or substantially improve user-facing interfaces with a product-specific visual direction, content hierarchy, and rendered critique.

    374 GitHub stars~2.3k tokensUpdated today
    Auto-check passed
  • Design Critique

    WrongStack/WrongStack

    A skill your agent uses to audit an interface that already exists and say precisely why it looks generated, templated, or unfinished — a scored rubric across composition, typography, color, states…

    374 GitHub stars~2.5k tokensUpdated today
    Auto-check passed
  • Mailbox Bridge

    WrongStack/WrongStack

    A skill your agent uses when external coding agents (Claude Code, Aider, custom scripts) need to participate in the project's shared WrongStack mailbox, or when a user asks to "expose the mailbox"…

    374 GitHub stars~2.6k tokensUpdated today
    Auto-check passed

Categories

Questions about Visual Regression

What does Visual Regression do?

Detect unintended UI changes with reproducible rendered screenshots and reviewed baselines. Visual Regression is an agent skill from WrongStack/WrongStack. Detect unintended UI changes with reproducible rendered screenshots and reviewed baselines.

When should I use Visual Regression?

Visual Regression fits situations like: adding screenshot checks; investigating layout/theme regressions; do not approve a changed image merely to make the test green.

How do I install Visual Regression in Claude Code?

Run `npx skills add WrongStack/WrongStack --skill visual-regression -a claude-code`. Or copy the skill folder (packages/core/skills/visual-regression in WrongStack/WrongStack) into .claude/skills/visual-regression in your project. Claude Code loads it when a task matches its description.

How do I install Visual Regression in Codex?

Run `npx skills add WrongStack/WrongStack --skill visual-regression -a codex`. Or copy the skill folder (packages/core/skills/visual-regression in WrongStack/WrongStack) into .agents/skills/visual-regression in your project. Codex loads it when a task matches its description.

Can I use Visual Regression in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add WrongStack/WrongStack --skill visual-regression -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/visual-regression, .gemini/skills/visual-regression, .github/skills/visual-regression and .opencode/skills/visual-regression in your project.

What does Visual Regression need to run?

SKILL.md names no scripts, command-line tools or credentials: Visual Regression is instructions for the agent only.

Does Visual Regression access the network?

SKILL.md names 2 domains. As links in the text: playwright.dev and storybook.js.org. This is read from the text; nothing was executed.

Is Visual Regression safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Visual Regression use?

Visual Regression is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Visual Regression use?

About 750 tokens (SKILL.md is roughly 3k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Visual Regression?

Skills that share tags, products or a category with Visual Regression: Review Local UI Screenshots (xiaocang/easydict_win32, 104 stars), Handsontable Visual Test Demos (handsontable/handsontable, 22k stars), Kc Screenshot (imran31415/kube-coder, 388 stars) and Visual QA (liangdabiao/Godogen, 126 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Visual Regression?

WrongStack (a GitHub organization) maintains it in WrongStack/WrongStack, which has 371 GitHub stars. The repository holds 100 skills in this directory. The repository was last updated on October 10, 2026.

Source: WrongStack/WrongStack on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.