Agent skill

Heuristic Evaluation

by Owl-Listener in Owl-Listener/designer-skills

Run an expert review against Nielsen's heuristics and domain criteria, with severity ratings.

MITAuto-check passedFrontend & Design

Install Heuristic Evaluation

skills CLI
$ npx skills add Owl-Listener/designer-skills --skill heuristic-evaluation -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install Owl-Listener/designer-skills heuristic-evaluation --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/Owl-Listener/designer-skills.git skills-src && mkdir -p .claude/skills && cp -r skills-src/prototyping-testing/skills/heuristic-evaluation .claude/skills/heuristic-evaluation && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
heuristic-evaluation
GitHub stars
2.9k
Used in
1 other repo
Token cost
~491 tokens
SKILL.md length
230 words
Files
1
Skills in repo
107
Repo updated
First seen
Licence
MIT

At a glance

Run an expert review against Nielsen's heuristics and domain criteria, with severity ratings.

  • Works in 10 steps: Visibility of system status — Users know… → Match real world — System speaks users'… → User control and freedom — Easy undo and… → …
  • You need findings without recruiting participants
  • SKILL.md covers What You Do, Nielsen's 10 Usability…, Evaluation Process and Issue Documentation, plus 2 more sections
  • Instructions only: no scripts, shell commands, URLs or credentials in SKILL.md

What it does

Heuristic Evaluation is an agent skill from Owl-Listener/designer-skills. Run an expert review against Nielsen's heuristics and domain criteria, with severity ratings. Use when you need findings without recruiting participants. For a facilitated team feedback session, use design-critique (design-ops).

Its SKILL.md is about 490 tokens, which your agent loads only when the skill is triggered. It is a single SKILL.md file with no bundled scripts.

It sits in Frontend & Design, covering UX design, Recruiting and HR and Design review and critique. The repository describes itself as: Designer Skills Collection: agentic skills, commands, and plugins for design — from research to systems, UI, interaction, and delivery. The licence is MIT.

When your agent uses it

  • You need findings without recruiting participants
  • Tasks that involve UX design
  • Tasks that involve Recruiting and HR

Example prompts

  • “/heuristic-evaluation”

Workflow steps

10 steps, taken from the first numbered list in SKILL.md.

  1. Visibility of system status — Users know what is happening
  2. Match real world — System speaks users' language
  3. User control and freedom — Easy undo and exit
  4. Consistency and standards — Follow conventions
  5. Error prevention — Prevent problems before they occur
  6. Recognition over recall — Make options visible
  7. Flexibility and efficiency — Shortcuts for experts
  8. Aesthetic and minimalist design — No irrelevant information
  9. Error recovery — Help users recognize and recover from errors
  10. Help and documentation — Provide assistance when needed

What it can do on your machine

Read from SKILL.md and the folder at commit 9a6930c. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    No scripts in the folder and no shell commands in SKILL.md.

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Heuristic Evaluation loads about 491 tokens when it runs. Until then it costs about 63 tokens; SKILL.md has 230 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~63
When it runs · the whole SKILL.md, loaded when a task matches
~491

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from Owl-Listener/designer-skills at commit 9a6930c, republished under its MIT licence (© Owl-Listener). 230 words, ~491 tokens.

Download SKILL.mdSave it as .claude/skills/heuristic-evaluation/SKILL.md (or your agent's skills folder).
name
heuristic-evaluation
description
Run an expert review against Nielsen's heuristics and domain criteria, with severity ratings. Use when you need findings without recruiting participants. For a facilitated team feedback session, use `design-critique` (design-ops).

Heuristic Evaluation

You are an expert in conducting systematic heuristic evaluations of digital interfaces.

What You Do

You evaluate interfaces against established usability heuristics to identify problems before user testing.

Nielsen's 10 Usability Heuristics

  1. Visibility of system status — Users know what is happening
  2. Match real world — System speaks users' language
  3. User control and freedom — Easy undo and exit
  4. Consistency and standards — Follow conventions
  5. Error prevention — Prevent problems before they occur
  6. Recognition over recall — Make options visible
  7. Flexibility and efficiency — Shortcuts for experts
  8. Aesthetic and minimalist design — No irrelevant information
  9. Error recovery — Help users recognize and recover from errors
  10. Help and documentation — Provide assistance when needed

Evaluation Process

  1. Define scope (which screens/flows to evaluate)
  2. Walk through as a new user
  3. Walk through as an experienced user
  4. Walk through each task flow
  5. Document each issue found
  6. Rate severity
  7. Compile and prioritize findings

Issue Documentation

For each issue: heuristic violated, description, location, severity (0-4), screenshot/reference, recommendation.

Severity Scale

  • 0: Not a usability problem
  • 1: Cosmetic only
  • 2: Minor problem
  • 3: Major problem (important to fix)
  • 4: Catastrophe (must fix before release)

Best Practices

  • Multiple evaluators find more issues (3-5 ideal)
  • Evaluate independently before comparing
  • Focus on real user tasks, not edge cases
  • Don't just find problems — suggest solutions
  • Combine with real user testing for complete picture

© Owl-Listener, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

Just SKILL.md in prototyping-testing/skills/heuristic-evaluation of Owl-Listener/designer-skills.

Open the folder on GitHubat commit 9a6930c

Used in 1 other repository

We found 2 copies of this SKILL.md (exact, near-identical or edited) in other folders, from 1 other GitHub owner. This page covers the copy in Owl-Listener/designer-skills, which our catalogue first saw on October 7, 2026.

Compare with similar skills

Heuristic Evaluation next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Heuristic Evaluation compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Heuristic Evaluation this skillOwl-Listener/designer-skills2.9k1 repos~491Automated safety check: PassMIT
Interface Design for Dashboards and Appsholaboss-ai/holaOS11k3 repos~6kAutomated safety check: PassMIT
Design Critiquegetcrew44/crew44356—~691Automated safety check: PassMIT
UX Surface Auditlobehub/lobehub83k—~4.9kAutomated safety check: PassCustom licence
Product Design Workflow BundleXiaomiMiMo/MiMo-Code14k—~721Automated safety check: PassMIT
Om UX Review PRgo-musicfox/go-musicfox2.6k1 repos~1.4kAutomated safety check: NotesGPL-3.0

Similar skills

  • Pushes an agent past generic defaults when designing dashboards, admin panels, SaaS apps and tools, with attention to structure, type, navigation and how data is shown.

    11k GitHub starsUsed in 3 repos~6k tokens
    Frontend & DesignAuto-check passed
  • Design Critique

    getcrew44/crew44

    Walks an agent through five review passes on a screen, flow or mockup, then returns ranked findings with a severity and a suggested fix for each one.

    356 GitHub stars~691 tokensUpdated 3 mo ago
    Frontend & DesignAuto-check passed
  • UX Surface Audit

    lobehub/lobehub

    Runs a repeatable UX review of one screen against the Designing Interfaces pattern language and LobeHub's ux checklists, using static, visual and dynamic layers.

    83k GitHub stars~4.9k tokensUpdated yesterday
    Frontend & DesignAuto-check passed
  • Product Design Workflow Bundle

    XiaomiMiMo/MiMo-Code

    Entry point to a bundle of product design workflows covering context, research, audits, ideation, URL or image to code, design QA and sharing a prototype.

    14k GitHub stars~721 tokensUpdated 5 days ago
    Frontend & DesignAuto-check passed
  • Om UX Review PR

    go-musicfox/go-musicfox

    Evidence-first design review of a PR's UI. An agent skill from go-musicfox/go-musicfox.

    2.6k GitHub starsUsed in 1 repo~1.4k tokens
    Frontend & DesignAuto-check: notes
  • Design Review

    jezweb/claude-skills

    Review a web app or page for visual design quality — layout, typography, spacing, colour, hierarchy, consistency, interaction patterns, and responsive behaviour.

    1.1k GitHub stars~2.1k tokensUpdated 3 days ago
    Frontend & DesignAuto-check passed

More from Owl-Listener/designer-skills

All 107 skills in this repo
  • A B Test Design

    Owl-Listener/designer-skills

    Design an A/B experiment — hypothesis, variants, primary metric, and sample size.

    2.9k GitHub starsUsed in 1 repo~472 tokens
    Auto-check passed
  • Accessibility Test Plan

    Owl-Listener/designer-skills

    Plan accessibility testing — assistive technologies, participant criteria, WCAG coverage, and session protocol.

    2.9k GitHub starsUsed in 1 repo~478 tokens
    Auto-check passed
  • Aesthetic Usability

    Owl-Listener/designer-skills

    Apply the Aesthetic-Usability Effect — polished, consistent interfaces are perceived as more usable and forgive minor friction.

    2.9k GitHub starsUsed in 1 repo~678 tokens
    Auto-check passed
  • Affinity Diagram

    Owl-Listener/designer-skills

    Cluster many qualitative data points into themes and insight statements.

    2.9k GitHub starsUsed in 1 repo~520 tokens
    Auto-check passed
  • Animation Principles

    Owl-Listener/designer-skills

    Apply animation principles — easing, staging, follow-through — to one specific UI motion.

    2.9k GitHub starsUsed in 1 repo~492 tokens
    Auto-check passed
  • Business Design

    Owl-Listener/designer-skills

    Read financials, map competitive landscapes, and argue design decisions in the language of value.

    2.9k GitHub starsUsed in 1 repo~1.1k tokens
    Auto-check passed

Questions about Heuristic Evaluation

What does Heuristic Evaluation do?

Run an expert review against Nielsen's heuristics and domain criteria, with severity ratings. Heuristic Evaluation is an agent skill from Owl-Listener/designer-skills. Run an expert review against Nielsen's heuristics and domain criteria, with severity ratings.

When should I use Heuristic Evaluation?

Heuristic Evaluation fits situations like: you need findings without recruiting participants; tasks that involve UX design; tasks that involve Recruiting and HR.

How do I install Heuristic Evaluation in Claude Code?

Run `npx skills add Owl-Listener/designer-skills --skill heuristic-evaluation -a claude-code`. Or copy the skill folder (prototyping-testing/skills/heuristic-evaluation in Owl-Listener/designer-skills) into .claude/skills/heuristic-evaluation in your project. Claude Code loads it when a task matches its description.

How do I install Heuristic Evaluation in Codex?

Run `npx skills add Owl-Listener/designer-skills --skill heuristic-evaluation -a codex`. Or copy the skill folder (prototyping-testing/skills/heuristic-evaluation in Owl-Listener/designer-skills) into .agents/skills/heuristic-evaluation in your project. Codex loads it when a task matches its description.

Can I use Heuristic Evaluation in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add Owl-Listener/designer-skills --skill heuristic-evaluation -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/heuristic-evaluation, .gemini/skills/heuristic-evaluation, .github/skills/heuristic-evaluation and .opencode/skills/heuristic-evaluation in your project.

What does Heuristic Evaluation need to run?

SKILL.md names no scripts, command-line tools or credentials: Heuristic Evaluation is instructions for the agent only.

Does Heuristic Evaluation access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is Heuristic Evaluation safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Heuristic Evaluation use?

Heuristic Evaluation is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Heuristic Evaluation use?

About 491 tokens (SKILL.md is roughly 2k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Heuristic Evaluation?

Skills that share tags, products or a category with Heuristic Evaluation: Interface Design for Dashboards and Apps (holaboss-ai/holaOS, 11k stars), Design Critique (getcrew44/crew44, 356 stars), UX Surface Audit (lobehub/lobehub, 83k stars) and Product Design Workflow Bundle (XiaomiMiMo/MiMo-Code, 14k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Heuristic Evaluation?

Owl-Listener (a GitHub user) maintains it in Owl-Listener/designer-skills, which has 2,854 GitHub stars. The repository holds 107 skills in this directory. The repository was last updated on September 5, 2026.

Source: Owl-Listener/designer-skills on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.