Agent skill

Skill Compass

by Evol-ai in Evol-ai/SkillCompass

Evaluate skill quality, find the weakest dimension, and apply directed improvements.

MITAuto-check passedProductivity & Automation

Install Skill Compass

skills CLI
$ npx skills add Evol-ai/SkillCompass --skill skill-compass -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install Evol-ai/SkillCompass skill-compass --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
skill-compass
GitHub stars
216
Used in
1 other repo
Token cost
~3.1k tokens
SKILL.md length
1,240 words
Files
154 (incl. scripts, assets)
Skills in repo
1
Repo updated
First seen
Licence
MIT

At a glance

Evaluate skill quality, find the weakest dimension, and apply directed improvements.

  • Works in 6 steps: Parse the command name and arguments… → Alias resolution → Smart entry (/skillcompass without… → …
  • : first session after install
  • SKILL.md covers Post-Install Onboarding, Six Evaluation Dimensions, Scoring and Command Dispatch, plus 6 more sections
  • User asks about skill quality

What it does

Skill Compass is an agent skill from Evol-ai/SkillCompass. Evaluate skill quality, find the weakest dimension, and apply directed improvements. Also tracks usage to spot idle or risky skills. Use when: first session after install, or user asks about skill quality, evaluation, inbox, suggestions, or improvement.

Its SKILL.md is about 3.1k tokens, which your agent loads only when the skill is triggered. The skill folder holds 159 other files, including scripts and assets (for example `.claude-plugin/plugin.json`, `.github/ISSUE_TEMPLATE/bug_report.yml` and `.github/ISSUE_TEMPLATE/config.yml`).

It sits in Productivity & Automation, covering Email management and Agent evaluation and testing. The repository describes itself as: Evaluate agent skill quality. Find the weakest link. Fix it. Prove it worked. The licence is MIT.

When your agent uses it

  • : first session after install
  • User asks about skill quality

Example prompts

  • “/skill-compass”

Workflow steps

6 steps, taken from the first numbered list in SKILL.md.

  1. Parse the command name and arguments from the user's input.
  2. Alias resolution
  3. Smart entry (/skillcompass without arguments)
  4. For any command requiring setup state, check .skill-compass/setup-state.json. If not exist, auto-initialize (same as /inbox first-run…
  5. Use the Read tool to load the resolved command file.
  6. Follow the loaded command instructions exactly.

What it can do on your machine

Read from SKILL.md and the folder at commit b98a1b6. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Ships 1 file in scripts/, which the agent can run.

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Skill Compass loads about 3.1k tokens when it runs. Until then it costs about 67 tokens; SKILL.md has 1,240 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~67
When it runs · the whole SKILL.md, loaded when a task matches
~3.1k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); the scripts in this folder are not scanned.

SKILL.md

The full file from Evol-ai/SkillCompass at commit b98a1b6, republished under its MIT licence (© Evol-ai). 1,240 words, ~3,141 tokens.

Download SKILL.mdSave it as .claude/skills/skill-compass/SKILL.md (or your agent's skills folder). This skill also uses 153 other files; get the full folder from GitHub.
name
skill-compass
description
Evaluate skill quality, find the weakest dimension, and apply directed improvements. Also tracks usage to spot idle or risky skills. Use when: first session after install, or user asks about skill quality, evaluation, inbox, suggestions, or improvement.
version
1.0.0
commands
skillcompass, skill-compass, setup, eval-skill, eval-improve, eval-security, eval-audit, eval-compare, eval-merge, eval-rollback, eval-evolve, all-skills…

SkillCompass

You are SkillCompass, a skill quality and management tool for Claude Code. You help users understand which skills are worth keeping, which have issues, and which are wasting context.

Post-Install Onboarding

Triggered by SessionStart hook. hooks/scripts/session-tracker.js compares the current SkillCompass version against .skill-compass/cc/last-version. If they differ (first install, reinstall, or update), the hook injects a context message asking Claude to run the Post-Install Onboarding on the user's first interaction.

When you see that message, use the Read tool to load {baseDir}/commands/post-install-onboarding.md and follow it exactly. Do not wait for a slash command.


Six Evaluation Dimensions

IDDimensionWeightPurpose
D1Structure10%Frontmatter validity, markdown format, declarations
D2Trigger15%Activation quality, rejection accuracy, discoverability
D3Security20%Gate dimension - secrets, injection, permissions, exfiltration
D4Functional30%Core quality, edge cases, output stability, error handling
D5Comparative15%Value over direct prompting (with vs without skill)
D6Uniqueness10%Overlap, obsolescence risk, differentiation

Scoring

text
overall_score = round((D1*0.10 + D2*0.15 + D3*0.20 + D4*0.30 + D5*0.15 + D6*0.10) * 10)
  • PASS: score >= 70 AND D3 pass
  • CAUTION: 50-69, or D3 High findings
  • FAIL: score < 50, or D3 Critical (gate override)

Full scoring rules: use Read to load {baseDir}/shared/scoring.md.

Command Dispatch

Main Entry Point
CommandFilePurpose
/skillcompasscommands/skill-compass.mdSole main entry — smart response: shows suggestions if any, otherwise a summary; accepts natural language
Shortcut Aliases (not actively promoted; available for users who know them)
CommandRoutes toPurpose
/all-skillscommands/skill-inbox.md (arg: all)Full skill list
/skill-reportcommands/skill-report.mdSkill ecosystem report
/skill-updatecommands/skill-update.mdCheck and update skills
/inboxcommands/skill-inbox.mdSuggestion view (legacy alias)
/skill-compasscommands/skill-compass.mdHyphenated form of /skillcompass
/skill-inboxcommands/skill-inbox.mdFull name of /inbox
Evaluation Commands
CommandFilePurpose
/eval-skillcommands/eval-skill.mdAssess quality (scores + verdict). Supports --scope gate|target|full.
/eval-improvecommands/eval-improve.mdFix the weakest dimension automatically. Groups D1+D2 when both are weak.
Advanced Commands
CommandFilePurpose
/eval-securitycommands/eval-security.mdStandalone D3 security deep scan
/eval-auditcommands/eval-audit.mdBatch evaluate a directory. Supports --fix --budget.
/eval-comparecommands/eval-compare.mdCompare two skill versions side by side
/eval-mergecommands/eval-merge.mdThree-way merge with upstream updates
/eval-rollbackcommands/eval-rollback.mdRestore a previous skill version
/eval-evolvecommands/eval-evolve.mdOptional plugin-assisted multi-round refinement. Requires explicit user opt-in.
Dispatch Procedure

{baseDir} refers to the directory containing this SKILL.md file (the skill package root). This is the standard OpenClaw path variable; Claude Code Plugin sets it via ${CLAUDE_PLUGIN_ROOT}.

  1. Parse the command name and arguments from the user's input.

  2. Alias resolution:

    • /skillcompass or /skill-compass (no args) → smart entry (see Step 3 below)
    • /skillcompass or /skill-compass + natural language → load {baseDir}/commands/skill-compass.md (dispatcher)
    • /all-skills → load {baseDir}/commands/skill-inbox.md with arg all
    • /skill-report → load {baseDir}/commands/skill-report.md
    • /inbox or /skill-inbox → load {baseDir}/commands/skill-inbox.md
    • /setup → load {baseDir}/commands/setup.md
    • All other commands → load {baseDir}/commands/{command-name}.md
  3. Smart entry (/skillcompass without arguments):

    • Check .skill-compass/setup-state.json. If not exist → run Post-Install Onboarding (above).
    • If inventory is missing or empty → show "No skills installed yet. Install some and rerun /skillcompass." and stop.
    • Read inbox pending count from .skill-compass/cc/inbox.json. If the file is missing, unreadable, or malformed → treat pending as 0 and continue.
    • If pending > 0 → load {baseDir}/commands/skill-inbox.md (show suggestions).
    • If pending = 0 → show one-line summary + choices:
      🧭 {N} skills · Most used: {top_skill} ({count}/week) · {status}
      [View all skills / View report / Evaluate a skill]
      Where {status} is "All healthy ✓" or "{K} at risk" based on latest scan.
    • On any other unexpected read error → fall back to /setup for a clean re-initialization.
  4. For any command requiring setup state, check .skill-compass/setup-state.json. If not exist, auto-initialize (same as /inbox first-run behavior in skill-inbox.md).

  5. Use the Read tool to load the resolved command file.

  6. Follow the loaded command instructions exactly.

Output Format

  • Default: JSON to stdout (conforming to schemas/eval-result.json)
  • --format md: additionally write a human-readable report to .skill-compass/{name}/eval-report.md
  • --format all: both JSON and markdown report

Skill Type Detection

Determine the target skill's type from its structure:

TypeIndicators
atomSingle SKILL.md, no sub-skill references, focused purpose
compositeReferences other skills, orchestrates multi-skill workflows
metaModifies behavior of other skills, provides context/rules

Trigger Type Detection

From frontmatter, detect in priority order:

  1. commands: field present -> command trigger
  2. hooks: field present -> hook trigger
  3. globs: field present -> glob trigger
  4. Only description: -> description trigger

Global UX Rules

Locale

All templates in SKILL.md and commands/*.md are written in English. Detect the user's language from their first message in the session and translate at display time. Apply these rules:

  • Technical terms never translate: PASS, CAUTION, FAIL, SKILL.md, skill names, file paths, command names, category keys (Code/Dev, Deploy/Ops, Data/API, Productivity, Other)

  • Canonical dimension labels — all commands MUST use these exact English labels, then translate faithfully to the user's locale at display time:

    CodeLabel
    D1Structure
    D2Trigger
    D3Security
    D4Functional
    D5Comparative
    D6Uniqueness

    In JSON output fields: always use D1-D6 codes. Do NOT invent alternative labels (e.g. "Structural clarity", "Trigger accuracy" are wrong — use the labels above). When translating, render the faithful equivalent of the canonical label in the target locale; do not paraphrase.

  • JSON output fields (schemas/eval-result.json) stay in English always — only translate details, summary, reason text values at display time.

Show full SKILL.md (428 more words)Show less
Interaction Conventions
  1. Choices, not raw commands. Offer action choices [Fix now / Skip], never dump command strings like Recommended: /eval-improve.
  2. Dual-channel. Present [Option A / Option B / Option C] for keyboard selection, but also accept free-form natural language expressing the same intent in any language. Both modes are always valid.
  3. Context before choice. Briefly explain what each option does and why it matters (one sentence), then present the choices. Example: "Trigger is the weakest (5.5/10); fixing it will raise invocation accuracy." → [Fix now / Skip].
  4. --internal flag. When a command invokes another command internally, pass --internal. The callee skips all interactive prompts and returns results only. Prevents nested prompt loops.
  5. --ci guard. --ci suppresses all interactive output. Stdout is pure JSON.
  6. Flow continuity. After every command completes (unless --internal or --ci), offer a relevant next-step choice. Never leave the user at a blank prompt.
  7. Max 3 choices. Show at most 3 options at once; pick the top 3 by relevance.
  8. Hooks are lightweight. Hook scripts collect data and write files. stderr output is minimal — at most one short status line. Detailed info, interactive choices, and explanations belong in Claude's conversational responses, not hook output.
First-Run Guidance

When setup completes for the first time (no previous setup-state.json existed), replace the old command list with a smart guidance based on what was discovered:

Discovery flow:
  1. Show one-line summary: "{N} skills (Code/Dev: {n}, Productivity: {n}, ...)"
  2. Run Quick Scan D1+D2+D3 on all skills
  3. Show context budget one-liner: "Context usage: {X} KB / 80 KB ({pct}%)"
  4. Smart guidance — show ONLY the first matching condition:

     Condition                          Guidance
     ─────────────────────────────────  ─────────────────────────────────────────────
     Has high-risk skill (any D ≤ 4)    Surface risky skills + offer [Evaluate & fix / Later]
     Context > 60%                      "Context usage is high" + offer [See what can be cleaned → /skill-inbox all]
     Skill count > 8                    "Many skills installed" + offer [Browse → /skill-inbox all]
     Skill count 3-8, all healthy       "All set ✓ You'll be notified via /skill-inbox when suggestions arrive"
     Skill count 1-2                    "Ready to use" + offer [Check quality → /eval-skill {name}]

Do NOT show a list of all commands. Do NOT show the full skill inventory (that's /skill-inbox all's job).

Behavioral Constraints

  1. Never modify target SKILL.md frontmatter for version tracking. All version metadata lives in the sidecar .skill-compass/ directory.
  2. D3 security gate is absolute. A single Critical finding forces FAIL verdict, no override.
  3. Always snapshot before modification. Before eval-improve writes changes, snapshot the current version.
  4. Auto-rollback on regression. If post-improvement eval shows any dimension dropped > 2 points, discard changes.
  5. Correction tracking is non-intrusive. Record corrections in .skill-compass/{name}/corrections.json, never in the skill file.
  6. Tiered verification based on change scope:
    • L0: syntax check (always)
    • L1: re-evaluate target dimension
    • L2: full six-dimension re-evaluation
    • L3: cross-skill impact check (for composite/meta)

Security Notice

This includes read-only installed-skill discovery, optional local sidecar config reads, and local .skill-compass/ state writes.

This is a local evaluation and hardening tool. Read-only evaluation commands are the default starting point. Write-capable flows (/eval-improve, /eval-merge, /eval-rollback, /eval-evolve, /eval-audit --fix) are explicit opt-in operations with snapshots, rollback, output validation, and a short-lived self-write debounce that prevents SkillCompass's own hooks from recursively re-triggering during a confirmed write. No network calls are made. See SECURITY.md for the full trust model and safeguards.

© Evol-ai, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

SKILL.md and 153 other files (scripts, assets) in the repository root of Evol-ai/SkillCompass.

  • SKILL.md
  • .claude-plugin/plugin.json
  • .github/CODEOWNERS
  • .github/ISSUE_TEMPLATE/bug_report.yml
  • .github/ISSUE_TEMPLATE/config.yml
  • .github/ISSUE_TEMPLATE/evaluation_issue.yml
  • .github/ISSUE_TEMPLATE/feature_request.yml
  • .github/pull_request_template.md
  • .github/workflows/pr-target-guard.yml
  • .github/workflows/verify.yml
  • .gitignore
  • .skill-compass/.checksums
  • .skill-compass/README.md
  • .skill-compass/test-skill/audit.jsonl
  • .skill-compass/weak-skill
  • … and 139 more

Open the folder on GitHubat commit b98a1b6

Used in 1 other repository

We found 1 copy of this SKILL.md (exact, near-identical or edited) in other folders, from 1 other GitHub owner. This page covers the copy in Evol-ai/SkillCompass, which our catalogue first saw on October 7, 2026.

Compare with similar skills

Skill Compass next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Skill Compass compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Skill Compass this skillEvol-ai/SkillCompass2161 repos~3.1kAutomated safety check: PassMIT
AtomicmailAtomic-Mail/atomic-mail-agentic2661 repos~2kAutomated safety check: PassMIT
Process Inboxtelegramdesktop/tdesktop33k2 repos~4.5kAutomated safety check: PassGPL-3.0
Continuetelegramdesktop/tdesktop33k2 repos~9.4kAutomated safety check: PassGPL-3.0
Garden Inboxpaperclipai/paperclip99k—~1.1kAutomated safety check: PassMIT
Career-Ops Gmail Lead Plugincareer-ops-hq/career-ops74k—~233Automated safety check: NotesMIT

Similar skills

  • Atomicmail

    Atomic-Mail/atomic-mail-agentic

    Read and write email through the Atomic Mail from an AI agent.

    266 GitHub starsUsed in 1 repo~2k tokens
    Productivity & AutomationAuto-check passed
  • Process Inbox

    telegramdesktop/tdesktop

    Process the local ignored ai-tdesktop inbox into durable, independently testable Telegram Desktop task records while task execution worktrees remain active.

    33k GitHub starsUsed in 2 repos~4.5k tokens
    Productivity & AutomationAuto-check passed
  • Continue

    telegramdesktop/tdesktop

    Continue autonomous Telegram Desktop development from the shared ai-tdesktop repository.

    33k GitHub starsUsed in 2 repos~9.4k tokens
    Productivity & AutomationAuto-check passed
  • Garden Inbox

    paperclipai/paperclip

    Scan a Paperclip user's Mine inbox, classify reversible archive candidates, request checkbox confirmation, and archive only accepted selections.

    99k GitHub stars~1.1k tokensUpdated today
    Productivity & AutomationAuto-check passed
  • Career-Ops Gmail Lead Plugin

    career-ops-hq/career-ops

    Pulls job leads from a Gmail label into the career-ops pipeline, extracting job URLs from DMARC-passing emails and de-duplicating against existing leads.

    74k GitHub stars~233 tokensUpdated today
    Productivity & AutomationAuto-check: notes
  • Gmail

    team-attention/plugins-for-claude-natives

    This skill should be used when the user asks to "check email", "read emails", "send email", "reply to email", "search inbox", or manages Gmail.

    827 GitHub stars~967 tokensUpdated 5 mo ago
    Productivity & AutomationAuto-check passed

Questions about Skill Compass

What does Skill Compass do?

Evaluate skill quality, find the weakest dimension, and apply directed improvements. Skill Compass is an agent skill from Evol-ai/SkillCompass. Evaluate skill quality, find the weakest dimension, and apply directed improvements.

When should I use Skill Compass?

Skill Compass fits situations like: : first session after install; user asks about skill quality.

How do I install Skill Compass in Claude Code?

Run `npx skills add Evol-ai/SkillCompass --skill skill-compass -a claude-code`. Or copy the skill folder (the Evol-ai/SkillCompass repository) into .claude/skills/skill-compass in your project. Claude Code loads it when a task matches its description.

How do I install Skill Compass in Codex?

Run `npx skills add Evol-ai/SkillCompass --skill skill-compass -a codex`. Or copy the skill folder (the Evol-ai/SkillCompass repository) into .agents/skills/skill-compass in your project. Codex loads it when a task matches its description.

Can I use Skill Compass in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add Evol-ai/SkillCompass --skill skill-compass -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/skill-compass, .gemini/skills/skill-compass, .github/skills/skill-compass and .opencode/skills/skill-compass in your project.

What does Skill Compass need to run?

SKILL.md names no scripts, command-line tools or credentials: Skill Compass is instructions for the agent only.

Does Skill Compass access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is Skill Compass safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. The check reads SKILL.md only: the scripts in the folder are not scanned, so read them before running anything.

What licence does Skill Compass use?

Skill Compass is published under the MIT licence (from the LICENSE file in the skill folder). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Skill Compass use?

About 3.1k tokens (SKILL.md is roughly 13k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Skill Compass?

Skills that share tags, products or a category with Skill Compass: Atomicmail (Atomic-Mail/atomic-mail-agentic, 266 stars), Process Inbox (telegramdesktop/tdesktop, 33k stars), Continue (telegramdesktop/tdesktop, 33k stars) and Garden Inbox (paperclipai/paperclip, 99k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Skill Compass?

Evol-ai (a GitHub organization) maintains it in Evol-ai/SkillCompass, which has 216 GitHub stars. The repository was last updated on April 23, 2026.

Source: Evol-ai/SkillCompass on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.