Agent skill

Eval

by jh941213 in jh941213/my-cc-harness

Evaluate code output on 4 axes (functionality/quality/originality/security) with scoring.

No licenceAuto-check: notesFrontend & Design

Install Eval

skills CLI
$ npx skills add jh941213/my-cc-harness --skill eval -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install jh941213/my-cc-harness eval --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/jh941213/my-cc-harness.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills_en/eval .claude/skills/eval && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
eval
GitHub stars
126
Token cost
~481 tokens
SKILL.md length
128 words
Files
1
Skills in repo
62
Repo updated
First seen
Licence
None found

At a glance

Evaluate code output on 4 axes (functionality/quality/originality/security) with scoring.

  • Works in 3 steps: Spawn the Evaluator agent → Review results → On CONDITIONAL/FAIL
  • Code evaluation
  • SKILL.md covers Process and pass@k idempotency test…
  • Instructions only: no scripts, shell commands, URLs or credentials in SKILL.md

What it does

Eval is an agent skill from jh941213/my-cc-harness. Evaluate code output on 4 axes (functionality/quality/originality/security) with scoring. Spawns an independent Evaluator agent. Triggers on: eval, evaluate, quality score, code evaluation. NOT for: writing code, implementation, review.

Its SKILL.md is about 480 tokens, which your agent loads only when the skill is triggered. It is a single SKILL.md file with no bundled scripts.

It sits in Frontend & Design, covering Accessibility.

When your agent uses it

  • Code evaluation
  • Tasks that involve Accessibility

Example prompts

  • “/eval”

Requirements

  • Pre-approved tools (allowed-tools): Read, Bash, Grep, Glob, Agent

Workflow steps

3 steps, taken from the step headings in SKILL.md.

  1. Spawn the Evaluator agent
  2. Review results
  3. On CONDITIONAL/FAIL

What it can do on your machine

Read from SKILL.md and the folder at commit e9210e2. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves these tools, so the agent can use them without asking each time:

    • Read
    • Bash
    • Grep
    • Glob
    • Agent

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    No scripts in the folder and no shell commands in SKILL.md (its code samples are bash).

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Eval loads about 481 tokens when it runs. Until then it costs about 60 tokens; SKILL.md has 128 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~60
When it runs · the whole SKILL.md, loaded when a task matches
~481

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check: notes

The automated check noted patterns worth knowing about, such as sudo or a known installer.

  • NotePre-approves every shell command (allowed-tools: Bash)SKILL.md
    allowed-tools: Read, Bash, Grep, Glob, Agent

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

Without a licence we can't republish the file, so here is its outline and opening line. It has 128 words (~481 tokens).

“Spawns an Evaluator agent separated from the Generator (implementer) for independent assessment.”

— opening of SKILL.md by jh941213
name
eval
allowed-tools
Read, Bash, Grep, Glob, Agent
user-invocable
true
disable-model-invocation
false

Read the full SKILL.md on GitHub

Files

Just SKILL.md in skills_en/eval of jh941213/my-cc-harness.

Open the folder on GitHubat commit e9210e2

Compare with similar skills

Eval next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Eval compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Eval this skilljh941213/my-cc-harness126—~481Automated safety check: NotesNone
Web Interface Guidelines Reviewervercel-labs/openreview1.7k98 repos~308Automated safety check: PassNone
Accessibility Reviewmarkmead/hyperui12k1 repos~1.1kAutomated safety check: PassMIT
Web Animation DesignbaptisteArno/typebot.io11k2 repos~2.7kAutomated safety check: PassCustom licence
Accessibility Fixeribelick/ui-skills9.5k4 repos~1.2kAutomated safety check: PassMIT
Wcag Audit PatternsvmDeshpande/ai-agent-automation17811 repos~610Automated safety check: PassApache-2.0

Similar skills

  • Web Interface Guidelines Reviewer

    vercel-labs/openreview

    Official

    Review UI code for Web Interface Guidelines compliance. Use when asked to "review my UI", "check accessibility", "audit design", "review UX", or "check my…

    1.7k GitHub starsUsed in 98 repos~308 tokens
    Frontend & DesignAuto-check passed
  • Accessibility Review

    markmead/hyperui

    Run a WCAG 2.1 AA accessibility audit on a design or page. An agent skill from markmead/hyperui.

    12k GitHub starsUsed in 1 repo~1.1k tokens
    Frontend & DesignAuto-check passed
  • Web Animation Design

    baptisteArno/typebot.io

    Guides easing, timing and animation choices for UI motion, based on a web animation course, and reviews existing animations in a before-and-after table.

    11k GitHub starsUsed in 2 repos~2.7k tokens
    Frontend & DesignAuto-check passed
  • Accessibility Fixer

    ibelick/ui-skills

    Audits and fixes HTML accessibility problems such as ARIA labels, keyboard navigation, focus management, contrast and form errors with minimal changes.

    9.5k GitHub starsUsed in 4 repos~1.2k tokens
    Frontend & DesignAuto-check passed
  • Wcag Audit Patterns

    vmDeshpande/ai-agent-automation

    Conduct WCAG 2.2 accessibility audits with automated testing, manual verification, and remediation guidance.

    178 GitHub starsUsed in 11 repos~610 tokens
    Frontend & DesignAuto-check passed
  • Baseline UI

    ibelick/ui-skills

    Applies a fixed set of UI rules for stack, components, interaction, animation, typography and layout, or reviews a file against them with concrete fixes.

    9.5k GitHub starsUsed in 8 repos~855 tokens
    Frontend & DesignAuto-check passed

More from jh941213/my-cc-harness

All 62 skills in this repo
  • API Design Principles

    jh941213/my-cc-harness

    REST 및 GraphQL API 설계 원칙 가이드. An agent skill from jh941213/my-cc-harness.

    126 GitHub starsUsed in 20 repos~3.4k tokens
    Auto-check passed
  • Tailwind Design System

    jh941213/my-cc-harness

    Build scalable design systems with Tailwind CSS, design tokens, component libraries, and responsive patterns.

    126 GitHub starsUsed in 10 repos~4.7k tokens
    Auto-check passed
  • Docs Architecture

    jh941213/my-cc-harness

    Generate/update architecture docs — ARCHITECTURE.md (codemap), architecture diagrams (C4 mermaid), ADRs (MADR), data model ERD.

    126 GitHub stars~923 tokensUpdated 2 mo ago
    Auto-check: notes
  • Auto Memory

    jh941213/my-cc-harness

    A skill your agent uses when starting substantial work in a repo (implementation, fixes, deploys, debugging), when first exploring a new repo, or when finishing work that produced reusable knowledge.

    126 GitHub stars~890 tokensUpdated 2 mo ago
    Auto-check passed
  • Async Python Patterns

    jh941213/my-cc-harness

    Python asyncio 및 async/await 패턴 가이드. An agent skill from jh941213/my-cc-harness.

    126 GitHub starsUsed in 14 repos~4.7k tokens
    Auto-check passed
  • Python Testing Patterns

    jh941213/my-cc-harness

    Implement comprehensive testing strategies with pytest, fixtures, mocking, and test-driven development.

    126 GitHub starsUsed in 17 repos~5.4k tokens
    Auto-check passed

Questions about Eval

What does Eval do?

Evaluate code output on 4 axes (functionality/quality/originality/security) with scoring. Eval is an agent skill from jh941213/my-cc-harness. Evaluate code output on 4 axes (functionality/quality/originality/security) with scoring.

When should I use Eval?

Eval fits situations like: code evaluation; tasks that involve Accessibility.

How do I install Eval in Claude Code?

Run `npx skills add jh941213/my-cc-harness --skill eval -a claude-code`. Or copy the skill folder (skills_en/eval in jh941213/my-cc-harness) into .claude/skills/eval in your project. Claude Code loads it when a task matches its description.

How do I install Eval in Codex?

Run `npx skills add jh941213/my-cc-harness --skill eval -a codex`. Or copy the skill folder (skills_en/eval in jh941213/my-cc-harness) into .agents/skills/eval in your project. Codex loads it when a task matches its description.

Can I use Eval in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add jh941213/my-cc-harness --skill eval -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/eval, .gemini/skills/eval, .github/skills/eval and .opencode/skills/eval in your project.

What does Eval need to run?

SKILL.md names no scripts, command-line tools or credentials: Eval is instructions for the agent only. Its frontmatter pre-approves these tools: Read, Bash, Grep, Glob, Agent.

Does Eval access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is Eval safe to install?

Our automated static check of SKILL.md found notes only (pre-approves every shell command (allowed-tools: bash)), nothing it rates as a warning. It is not a guarantee. Review the folder before installing.

What licence does Eval use?

No licence was found for Eval or its repository. Without one, default copyright applies: ask the author before reusing or redistributing it.

How many tokens does Eval use?

About 481 tokens (SKILL.md is roughly 1.9k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Eval?

Skills that share tags, products or a category with Eval: Web Interface Guidelines Reviewer (vercel-labs/openreview, 1.7k stars), Accessibility Review (markmead/hyperui, 12k stars), Web Animation Design (baptisteArno/typebot.io, 11k stars) and Accessibility Fixer (ibelick/ui-skills, 9.5k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Eval?

jh941213 (a GitHub user) maintains it in jh941213/my-cc-harness, which has 126 GitHub stars. The repository holds 62 skills in this directory. The repository was last updated on August 3, 2026.

Source: jh941213/my-cc-harness on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.