Agent skill

Rules Eval

by athola in athola/claude-night-market

Evaluate Claude Code rules in .claude/rules/. An agent skill from athola/claude-night-market.

MITAuto-check passed

Install Rules Eval

skills CLI
$ npx skills add athola/claude-night-market --skill rules-eval -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install athola/claude-night-market rules-eval --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/athola/claude-night-market.git skills-src && mkdir -p .claude/skills && cp -r skills-src/plugins/abstract/skills/rules-eval .claude/skills/rules-eval && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
rules-eval
GitHub stars
342
Token cost
~919 tokens
SKILL.md length
303 words
Files
5
Skills in repo
160
Repo updated
First seen
Licence
MIT

At a glance

Evaluate Claude Code rules in .claude/rules/. An agent skill from athola/claude-night-market.

  • Works in 6 steps: Scan .claude/rules/ for all .md files… → Validate YAML frontmatter syntax and… → Analyze glob patterns for correctness… → …
  • SKILL.md covers When NOT To Use, Overview, Quick Start and Evaluation Workflow, plus 3 more sections
  • Instructions only: no scripts, shell commands, URLs or credentials in SKILL.md

What it does

Rules Eval is an agent skill from athola/claude-night-market. Evaluate Claude Code rules in .claude/rules/. Use for frontmatter, globs, and quality audits.

Its SKILL.md is about 920 tokens, which your agent loads only when the skill is triggered. The skill folder holds 5 other files (for example `modules/content-quality-metrics.md`, `modules/frontmatter-validation.md` and `modules/glob-pattern-analysis.md`).

The repository describes itself as: 23 Claude Code plugins: TDD enforcement hooks, git/PR workflows, spec-driven development, code review, project lifecycle, fix-from-error, maintenance automation, context… The licence is MIT.

Example prompts

  • “/rules-eval”

Workflow steps

6 steps, taken from the first numbered list in SKILL.md.

  1. Scan .claude/rules/ for all .md files (including subdirectories)
  2. Validate YAML frontmatter syntax and fields
  3. Analyze glob patterns for correctness and specificity
  4. Assess content quality (actionable, concise, non-conflicting)
  5. Check organization (naming, structure, symlinks)
  6. Measure token efficiency and redundancy

What it can do on your machine

Read from SKILL.md and the folder at commit 9f3eb00. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    No scripts in the folder and no shell commands in SKILL.md (its code samples are bash).

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Rules Eval loads about 919 tokens when it runs. Until then it costs about 26 tokens; SKILL.md has 303 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~26
When it runs · the whole SKILL.md, loaded when a task matches
~919

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from athola/claude-night-market at commit 9f3eb00, republished under its MIT licence (© athola). 303 words, ~919 tokens.

Download SKILL.mdSave it as .claude/skills/rules-eval/SKILL.md (or your agent's skills folder). This skill also uses 4 other files; get the full folder from GitHub.
name
rules-eval
description
Evaluate Claude Code rules in .claude/rules/. Use for frontmatter, globs, and quality audits.
alwaysApply
false
category
rule-management
tags
evaluation, rules, validation, quality-assurance, glob-patterns, frontmatter
dependencies
skills-eval
provides.infrastructure
rules-evaluation, frontmatter-validation, glob-pattern-analysis
provides.patterns
rule-auditing, content-quality, organization-analysis
estimated_tokens
400
evaluation_criteria.frontmatter_validity
25
evaluation_criteria.glob_pattern_quality
20
evaluation_criteria.content_quality
25

Rules Evaluation Framework

When NOT To Use

  • Evaluating skills (use abstract:skills-eval)
  • Evaluating hooks (use abstract:hooks-eval)
  • Writing a new rule (use hookify:writing-rules)

Overview

This skill evaluates Claude Code rules in .claude/rules/ directories against quality standards. It validates YAML frontmatter, glob pattern syntax, content quality, and directory organization. Rules files support path-scoped conditional loading via paths frontmatter and unconditional rules (no paths field).

Key validations: YAML syntax errors, unquoted glob patterns, Cursor-specific fields (alwaysApply, globs), overly broad patterns, content verbosity, and naming conventions.

Quick Start

bash
# Evaluate rules in current project
/rules-eval

# Evaluate specific directory
/rules-eval .claude/rules/

# Detailed analysis with recommendations
/rules-eval --detailed

Evaluation Workflow

  1. Scan .claude/rules/ for all .md files (including subdirectories)
  2. Validate YAML frontmatter syntax and fields
  3. Analyze glob patterns for correctness and specificity
  4. Assess content quality (actionable, concise, non-conflicting)
  5. Check organization (naming, structure, symlinks)
  6. Measure token efficiency and redundancy

Scoring

CategoryPointsFocus
Frontmatter Validity25YAML syntax, required fields, correct field names
Glob Pattern Quality20Syntax, specificity, quoting
Content Quality25Actionable, concise, non-conflicting
Organization15Naming, structure, symlink usage
Token Efficiency15Rule size, redundancy detection
ScoreLevel
91-100Excellent - Production-ready
76-90Good - Minor improvements possible
51-75Basic - Needs optimization
26-50Below Standards - Significant issues
0-25Critical - Invalid or broken rules

Resources

Skill-Specific Modules
  • Frontmatter Validation: See modules/frontmatter-validation.md
  • Glob Pattern Analysis: See modules/glob-pattern-analysis.md
  • Content Quality Metrics: See modules/content-quality-metrics.md
  • Organization Patterns: See modules/organization-patterns.md
Tools
  • Rules Validator: scripts/rules_validator.py
  • abstract:skills-eval - Skill evaluation framework
  • abstract:hooks-eval - Hook evaluation framework

Exit Criteria

  • Every .md file under .claude/rules/ (including subdirectories) receives a quality score (0-100) with per-category breakdown across the five dimensions.
  • YAML frontmatter syntax errors and unquoted glob patterns are listed as blocking findings before any score is reported.
  • Rules using Cursor-specific fields (alwaysApply, globs) are flagged as non-compliant with the expected Claude Code schema.
  • Any rule scoring below 26 (Critical tier) is reported with at least one concrete corrective action, not just a score.

© athola, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

SKILL.md and 4 other files in plugins/abstract/skills/rules-eval of athola/claude-night-market.

  • SKILL.md
  • modules/content-quality-metrics.md
  • modules/frontmatter-validation.md
  • modules/glob-pattern-analysis.md
  • modules/organization-patterns.md

Open the folder on GitHubat commit 9f3eb00

Compare with similar skills

Rules Eval next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Rules Eval compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Rules Eval this skillathola/claude-night-market342—~919Automated safety check: PassMIT
Eval-Driven Development Harnessaffaan-m/ECC274k—~1.5kAutomated safety check: PassMIT
Evalalirezarezvani/claude-skills28k1 repos~618Automated safety check: PassMIT
Eval Harnessaffaan-m/ECC274k—~2.2kAutomated safety check: PassMIT
Eval Harnessaffaan-m/ECC274k1 repos~1.7kAutomated safety check: PassMIT
OmniRoute CLI Evalsdiegosouzapw/OmniRoute74k—~1.3kAutomated safety check: PassMIT

Similar skills

  • Sets up eval-driven development for Claude Code workflows: capability and regression evals, three grader types and pass@k reliability metrics.

    274k GitHub stars~1.5k tokensUpdated 2 days ago
    Agent WorkflowsAuto-check passed
  • Eval

    alirezarezvani/claude-skills

    Evaluate and rank agent results by metric or LLM judge for an AgentHub session.

    28k GitHub starsUsed in 1 repo~618 tokens
    AI & LLM EngineeringAuto-check passed
  • Eval Harness

    affaan-m/ECC

    Eval-driven development (EDD) framework for AI coding sessions — define capability and regression evals before coding, grade with code-based, model-based, rule, or human graders, and track pass@k…

    274k GitHub stars~2.2k tokensUpdated 2 days ago
    AI & LLM EngineeringAuto-check passed
  • Eval Harness

    affaan-m/ECC

    Eval-driven development (EDD) ilkelerini uygulayan Claude Code oturumları için formal değerlendirme çerçevesi

    274k GitHub starsUsed in 1 repo~1.7k tokens
    AI & LLM EngineeringAuto-check passed
  • OmniRoute CLI Evals

    diegosouzapw/OmniRoute

    Creates and runs LLM evaluation suites from the omniroute CLI, follows live runs, shows scorecards, compares models and ties eval runs into CI.

    74k GitHub stars~1.3k tokensUpdated today
    AI & LLM EngineeringAuto-check passed
  • Arize Evaluator

    github/awesome-copilot

    Official

    Handles LLM-as-judge evaluation workflows on Arize including creating/updating evaluators, running evaluations on spans or experiments, managing tasks, trigger-run operations, column mapping, and…

    40k GitHub starsUsed in 2 repos~8.1k tokens
    AI & LLM EngineeringAuto-check: notes

More from athola/claude-night-market

All 160 skills in this repo
  • Night Market Diagnostics Toolkit

    athola/claude-night-market

    Run and interpret repo diagnostic scripts (ratchets, validators, token stats).

    342 GitHub stars~3.4k tokensUpdated yesterday
    Auto-check passed
  • Skills Eval

    athola/claude-night-market

    Evaluate Claude skill quality through auditing. An agent skill from athola/claude-night-market.

    342 GitHub stars~1.6k tokensUpdated yesterday
    Auto-check passed
  • Agent Teams

    athola/claude-night-market

    Coordinates Claude agent teams via filesystem protocol. An agent skill from athola/claude-night-market.

    342 GitHub stars~2.5k tokensUpdated yesterday
    Auto-check passed
  • Delegation Core

    athola/claude-night-market

    Delegates execution to eight CLIs (Gemini, Qwen, MiniMax, GLM, Muse, Codex, OpenCode, Glimmer).

    342 GitHub stars~2.5k tokensUpdated yesterday
    Auto-check passed
  • Elegant Code

    athola/claude-night-market

    Guide minimal code via a decision ladder with full safety, edge, and negative-case coverage.

    342 GitHub stars~2.1k tokensUpdated yesterday
    Auto-check passed
  • Skill Library Mission

    athola/claude-night-market

    Build a project skill library in .claude/skills/ via discovery, parallel authoring, and review.

    342 GitHub stars~1.6k tokensUpdated yesterday
    Auto-check passed

Questions about Rules Eval

What does Rules Eval do?

Evaluate Claude Code rules in .claude/rules/. An agent skill from athola/claude-night-market. Rules Eval is an agent skill from athola/claude-night-market.claude/rules/.

How do I install Rules Eval in Claude Code?

Run `npx skills add athola/claude-night-market --skill rules-eval -a claude-code`. Or copy the skill folder (plugins/abstract/skills/rules-eval in athola/claude-night-market) into .claude/skills/rules-eval in your project. Claude Code loads it when a task matches its description.

How do I install Rules Eval in Codex?

Run `npx skills add athola/claude-night-market --skill rules-eval -a codex`. Or copy the skill folder (plugins/abstract/skills/rules-eval in athola/claude-night-market) into .agents/skills/rules-eval in your project. Codex loads it when a task matches its description.

Can I use Rules Eval in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add athola/claude-night-market --skill rules-eval -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/rules-eval, .gemini/skills/rules-eval, .github/skills/rules-eval and .opencode/skills/rules-eval in your project.

What does Rules Eval need to run?

SKILL.md names no scripts, command-line tools or credentials: Rules Eval is instructions for the agent only.

Does Rules Eval access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is Rules Eval safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Rules Eval use?

Rules Eval is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Rules Eval use?

About 919 tokens (SKILL.md is roughly 3.7k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Rules Eval?

Skills that share tags, products or a category with Rules Eval: Eval-Driven Development Harness (affaan-m/ECC, 274k stars), Eval (alirezarezvani/claude-skills, 28k stars), Eval Harness (affaan-m/ECC, 274k stars) and Eval Harness (affaan-m/ECC, 274k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Rules Eval?

athola (a GitHub user) maintains it in athola/claude-night-market, which has 342 GitHub stars. The repository holds 160 skills in this directory. The repository was last updated on October 6, 2026.

Source: athola/claude-night-market on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.