Agent skill

Yida Skill Evaluator

by openyida in openyida/openyida

评测指定的 OpenYida 技能质量。输入技能名称,自动执行全链路闭环评测 (静态校验 → 路由测试 → 安全合规 → 覆盖度 → 多维评分 → 准出门槛 → 优化建议), 生成评测报告和改进建议。不要触发本技能来执行其他 OpenYida 开发任务。

MITAuto-check passedTesting & QA

Install Yida Skill Evaluator

skills CLI
$ npx skills add openyida/openyida --skill yida-skill-evaluator -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install openyida/openyida yida-skill-evaluator --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/openyida/openyida.git skills-src && mkdir -p .claude/skills && cp -r skills-src/yida-skills/skills/yida-skill-evaluator .claude/skills/yida-skill-evaluator && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
yida-skill-evaluator
GitHub stars
220
Token cost
~830 tokens
SKILL.md length
163 words
Files
1
Skills in repo
58
Repo updated
First seen
Licence
MIT

At a glance

评测指定的 OpenYida 技能质量。输入技能名称,自动执行全链路闭环评测 (静态校验 → 路由测试 → 安全合规 → 覆盖度 → 多维评分 → 准出门槛 → 优化建议), 生成评测报告和改进建议。不要触发本技能来执行其他 OpenYida 开发任务。

  • Works in 5 steps: 快速评测(默认,无副作用) → 解读结果 → 深度评测(可选,用户要求时) → …
  • Testing & QA work in your project
  • SKILL.md covers 触发条件, 前置条件, 工作流 and 可用评测模式, plus 1 more section
  • Calls node and npm

What it does

Yida Skill Evaluator is an agent skill from openyida/openyida. 评测指定的 OpenYida 技能质量。输入技能名称,自动执行全链路闭环评测 (静态校验 → 路由测试 → 安全合规 → 覆盖度 → 多维评分 → 准出门槛 → 优化建议), 生成评测报告和改进建议。不要触发本技能来执行其他 OpenYida 开发任务。

Its SKILL.md is about 830 tokens, which your agent loads only when the skill is triggered. It is a single SKILL.md file with no bundled scripts.

It sits in Testing & QA. The repository describes itself as: Your own personal YiDA AI assistant! The licence is MIT.

When your agent uses it

  • Testing & QA work in your project

Example prompts

  • “/yida-skill-evaluator”

Requirements

  • Node.js

Workflow steps

5 steps, taken from the step headings in SKILL.md.

  1. 快速评测(默认,无副作用)
  2. 解读结果
  3. 深度评测(可选,用户要求时)
  4. Web 控制台(可选)
  5. 输出评测报告

What it can do on your machine

Read from SKILL.md and the folder at commit 3dd4693. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Shell commands in SKILL.md call:

    • node
    • npm

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md. Its commands use npm, which can reach the network depending on how they are called.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Yida Skill Evaluator loads about 830 tokens when it runs. Until then it costs about 37 tokens; SKILL.md has 163 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~37
When it runs · the whole SKILL.md, loaded when a task matches
~830

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from openyida/openyida at commit 3dd4693, republished under its MIT licence (© openyida). 163 words, ~830 tokens.

Download SKILL.mdSave it as .claude/skills/yida-skill-evaluator/SKILL.md (or your agent's skills folder).
name
yida-skill-evaluator
description
评测指定的 OpenYida 技能质量。输入技能名称,自动执行全链路闭环评测 (静态校验 → 路由测试 → 安全合规 → 覆盖度 → 多维评分 → 准出门槛 → 优化建议), 生成评测报告和改进建议。不要触发本技能来执行其他 OpenYida 开发任务。
triggers
评测技能, 测试技能, skill eval, 评估技能质量, 跑一下评测, 技能打分, 检查技能, 技能准出, 评测 xx 技能, run eval

触发条件

当用户的意图是「评测/测试/检查某个 openyida 技能的质量」时触发。典型表述:

  • "评测一下 yida-dashboard 技能"
  • "帮我跑一下 skill eval"
  • "检查这个技能的质量"
  • "这个技能能不能准出?"
  • "评估 yida-report 技能"
  • "测试下路由准确率"

前置条件

  1. 确认目标技能名称(如 yida-dashboard)。若用户未指定,列出可用技能供选择:
    bash
    ls yida-skills/skills/
  2. 确认 Node.js 版本 ≥ 18:
    bash
    node --version
  3. 确认依赖已安装:
    bash
    npm install

工作流

Step 1 — 快速评测(默认,无副作用)

大多数情况使用 pipeline 模式,它会自动串联所有评测步骤:

bash
node scripts/eval/pipeline.js --skill <技能名>

该命令自动执行以下全部步骤:

  1. 静态校验(文档规范性 + 可维护性)
  2. 路由测试(命中率 + 混淆对)
  3. 安全合规检查
  4. 覆盖度分析
  5. 10 维度综合评分 + 加权总分
  6. 准出门槛判定 + 自动优化建议

产物目录:project/.cache/eval/pipeline/<run-id>/

读取结果:

bash
cat project/.cache/eval/pipeline/*/pipeline-report.json | tail -1
Step 2 — 解读结果

从 pipeline-report.json 中提取关键信息,向用户报告:

  1. Pipeline 状态:status 字段(pass / fail / warn)
  2. 各步骤结果:steps[] 数组的 step / status / score / detail
  3. 总分:scorecard.overall(0-100 分)
  4. 准出判定:scorecard.gate(pass = 可准出,fail = 不可准出)
  5. 硬门槛详情:scorecard.hardGates 中每项的状态
  6. 优化建议:suggestions[] 中按 priority 排序的改进项
Step 3 — 深度评测(可选,用户要求时)

如果用户要求更深度的评测(如 A/B 对比、JUnit 报告):

bash
# A/B 基线对比(with_skill vs without_skill)
node scripts/eval/runner.js --mode baseline --skill <技能名> --format junit

# 单独维度评测
node scripts/eval/runner.js --mode doc-quality --skill <技能名>
node scripts/eval/runner.js --mode comprehensive --skill <技能名>
Step 4 — Web 控制台(可选)

如果用户希望在浏览器中查看评测结果:

bash
npm run eval:dashboard

然后打开 http://127.0.0.1:4500 查看控制台。

Step 5 — 输出评测报告

向用户呈现结构化报告:

## 技能评测报告:<技能名>

### Pipeline 状态:<PASS/FAIL/WARN>

| 步骤 | 状态 | 分数 | 详情 |
|------|------|------|------|
| 静态校验 | ✔/✗ | xx | ... |
| 路由测试 | ✔/✗ | xx% | ... |
| 安全合规 | ✔/✗ | xx | ... |
| 覆盖度 | ✔/✗ | xx% | ... |
| 综合评分 | ✔/✗ | xx/100 | ... |
| 准出判定 | ✔/✗ | - | ... |

### 评分卡(10 维度)

| 维度 | 得分 | 权重 | 加权得分 |
|------|------|------|----------|
| 规范性 | xx | 10% | x.x |
| 可维护性 | xx | 5% | x.x |
| 路由准确率 | xx | 15% | x.x |
| ... | ... | ... | ... |
| **总分** | **xx** | | |

### 准出门槛

| 门槛 | 要求 | 实际 | 状态 |
|------|------|------|------|
| 触发准确率 | ≥ 85% | xx% | ✔/✗ |
| 步骤完成率 | = 100% | xx% | ✔/✗ |
| 功能测试通过率 | ≥ 95% | xx% | ✔/✗ |
| 输出格式正确率 | ≥ 85% | xx% | ✔/✗ |
| 安全合规 | 0 failures | xx | ✔/✗ |

### 优化建议(前 3 条)

1. [blocker/critical/high] ...
2. ...
3. ...

可用评测模式

模式命令说明
pipeline--mode pipeline全自动闭环(推荐)
doc-quality--mode doc-quality文档规范性
routing--mode routing路由准确率
safety--mode safety安全合规
coverage--mode coverage覆盖度
comprehensive--mode comprehensive10 维度评分
baseline--mode baselineA/B 基线对比
e2e--mode e2e端到端基线(需 OPENYIDA_E2E=1)
generate--mode generate真实生成(需 OPENYIDA_E2E=1)

WHEN NOT

  • 不处理应用搭建、页面开发、数据管理等开发任务
  • 不处理非 openyida 技能的评测
  • 不修改技能文件——仅读取和评测(除非用户明确要求根据建议修改)
  • 不执行需要真实宜搭资源的评测(e2e / generate),除非用户明确要求
  • 不与 yida-dashboard、yida-report、yida-create-process 等搭建类技能混淆

© openyida, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

Just SKILL.md in yida-skills/skills/yida-skill-evaluator of openyida/openyida.

Open the folder on GitHubat commit 3dd4693

Compare with similar skills

Yida Skill Evaluator next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Yida Skill Evaluator compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Yida Skill Evaluator this skillopenyida/openyida220—~830Automated safety check: PassMIT
Web Application Testinganthropics/skills180k51 repos~966Automated safety check: PassApache-2.0
Diagnosing Bugsfossasia/eventyay-interpretation1.6k32 repos~2.1kAutomated safety check: PassApache-2.0
TDDpietheinstrengholt/rssmonster56430 repos~906Automated safety check: PassMIT
TDD WorkflowhellangleZ/burn-in-cceverywhere-ralph11211 repos~2.4kAutomated safety check: PassNone
TDDsanity-io/sanity6.4k20 repos~1kAutomated safety check: PassMIT

Similar skills

  • Web Application Testing

    anthropics/skills

    Official

    Tests local web applications with Python Playwright scripts, checking frontend behavior, capturing screenshots and reading browser console logs.

    180k GitHub starsUsed in 51 repos~966 tokens
    Testing & QAAuto-check passed
  • Diagnosing Bugs

    fossasia/eventyay-interpretation

    Diagnosis loop for hard bugs and performance regressions. An agent skill from fossasia/eventyay-interpretation.

    1.6k GitHub starsUsed in 32 repos~2.1k tokens
    Testing & QAAuto-check passed
  • TDD

    pietheinstrengholt/rssmonster

    Test-driven development. An agent skill from pietheinstrengholt/rssmonster.

    564 GitHub starsUsed in 30 repos~906 tokens
    Testing & QAAuto-check passed
  • TDD Workflow

    hellangleZ/burn-in-cceverywhere-ralph

    A skill your agent uses when writing new features, fixing bugs, or refactoring code.

    112 GitHub starsUsed in 11 repos~2.4k tokens
    Testing & QAAuto-check passed
  • TDD

    sanity-io/sanity

    Official

    Test-driven development with red-green-refactor loop. An agent skill from sanity-io/sanity.

    6.4k GitHub starsUsed in 20 repos~1k tokens
    Testing & QAAuto-check passed
  • Context Driven Development

    Ibrahim-3d/orchestrator-supaconductor

    A skill your agent uses when working with Conductor's context-driven development methodology, managing project context artifacts, or understanding the relationship between product.md, tech-stack.md…

    380 GitHub starsUsed in 9 repos~2.9k tokens
    Testing & QAAuto-check passed

More from openyida/openyida

All 58 skills in this repo
  • Yida Canvas Custom Page

    openyida/openyida

    宜搭自定义页面开发规范,使用 YidaCodeCanvas 组件实现现代 React18 自定义页面。用于官网、看板、工作台、列表、详情、门户壳、可视化、hooks 交互、表单入口,以及需要门户组件、数据管理视图、成员、部门或上传组件的场景;发布层自动注入 yida/utils window 桥。

    220 GitHub stars~4.3k tokensUpdated 8 days ago
    Auto-check passed
  • Openyida Release

    openyida/openyida

    快速发布 OpenYida npm 正式版或测试版。用户说“发布正式版”“发布测试版”“发 beta”或要求按当天日期生成版本 tag 时使用;不用于宜搭应用或自定义页面发布。

    220 GitHub stars~520 tokensUpdated 8 days ago
    Auto-check passed
  • Yida Design Plan

    openyida/openyida

    Plan 模式的视觉设计分支。基于需求选择视觉方向,再结合业务页面规划维护 build-plan.json 的 visualStyle,供 CLI 派生 design.md。

    220 GitHub stars~756 tokensUpdated 8 days ago
    Auto-check passed
  • Prevent stale local OpenYida custom page source from overwriting live designer edits.

    220 GitHub stars~1.2k tokensUpdated 8 days ago
    Auto-check passed
  • Codemap

    openyida/openyida

    Generate, update, or drift-check agent-facing CodeMaps as progressive code terrain indexes for projects, features, capabilities, functions, modules, or bug chains.

    220 GitHub stars~1k tokensUpdated 8 days ago
    Auto-check passed
  • Sdd Riper One Light

    openyida/openyida

    面向 GPT-5.5 等强模型和熟练用户的轻量 AI Agent Harness / checkpoint-driven coding skill。默认用户已经把任务切到基本可执行的最小混沌单元;模型自行分解、探索与推进,人类通过最终目标、最小 spec、复述、checkpoint、证据验证与回写来低干扰控盘。

    220 GitHub stars~1.5k tokensUpdated 8 days ago
    Auto-check passed

Categories

Questions about Yida Skill Evaluator

What does Yida Skill Evaluator do?

评测指定的 OpenYida 技能质量。输入技能名称,自动执行全链路闭环评测 (静态校验 → 路由测试 → 安全合规 → 覆盖度 → 多维评分 → 准出门槛 → 优化建议), 生成评测报告和改进建议。不要触发本技能来执行其他 OpenYida 开发任务。. Yida Skill Evaluator is an agent skill from openyida/openyida.

When should I use Yida Skill Evaluator?

Yida Skill Evaluator fits situations like: testing & QA work in your project.

How do I install Yida Skill Evaluator in Claude Code?

Run `npx skills add openyida/openyida --skill yida-skill-evaluator -a claude-code`. Or copy the skill folder (yida-skills/skills/yida-skill-evaluator in openyida/openyida) into .claude/skills/yida-skill-evaluator in your project. Claude Code loads it when a task matches its description.

How do I install Yida Skill Evaluator in Codex?

Run `npx skills add openyida/openyida --skill yida-skill-evaluator -a codex`. Or copy the skill folder (yida-skills/skills/yida-skill-evaluator in openyida/openyida) into .agents/skills/yida-skill-evaluator in your project. Codex loads it when a task matches its description.

Can I use Yida Skill Evaluator in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add openyida/openyida --skill yida-skill-evaluator -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/yida-skill-evaluator, .gemini/skills/yida-skill-evaluator, .github/skills/yida-skill-evaluator and .opencode/skills/yida-skill-evaluator in your project.

What does Yida Skill Evaluator need to run?

Going by SKILL.md and its folder, Yida Skill Evaluator needs the command-line tools its instructions call (node and npm). Our summary lists: Node.js.

Does Yida Skill Evaluator access the network?

SKILL.md contains no URLs. Its commands use npm, which can reach the network depending on how they are called. This is read from the text; nothing was executed.

Is Yida Skill Evaluator safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Yida Skill Evaluator use?

Yida Skill Evaluator is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Yida Skill Evaluator use?

About 830 tokens (SKILL.md is roughly 3.3k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Yida Skill Evaluator?

Skills that share tags, products or a category with Yida Skill Evaluator: Web Application Testing (anthropics/skills, 180k stars), Diagnosing Bugs (fossasia/eventyay-interpretation, 1.6k stars), TDD (pietheinstrengholt/rssmonster, 564 stars) and TDD Workflow (hellangleZ/burn-in-cceverywhere-ralph, 112 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Yida Skill Evaluator?

openyida (a GitHub organization) maintains it in openyida/openyida, which has 220 GitHub stars. The repository holds 58 skills in this directory. The repository was last updated on September 30, 2026.

Source: openyida/openyida on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.