Agent skill

Bernstein Quality

by sipyourdrink-ltd in sipyourdrink-ltd/bernstein

Show quality metrics for Bernstein runs - success rates per model, lint/test pass rates, completion time distributions.

Apache-2.0Auto-check passedDevelopment

Install Bernstein Quality

skills CLI
$ npx skills add sipyourdrink-ltd/bernstein --skill bernstein-quality -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install sipyourdrink-ltd/bernstein bernstein-quality --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/sipyourdrink-ltd/bernstein.git skills-src && mkdir -p .claude/skills && cp -r skills-src/packages/cursor-plugin/skills/bernstein-quality .claude/skills/bernstein-quality && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
bernstein-quality
GitHub stars
1.4k
Token cost
~388 tokens
SKILL.md length
102 words
Files
2 (incl. scripts)
Skills in repo
9
Repo updated
First seen
Licence
Apache-2.0

At a glance

Show quality metrics for Bernstein runs - success rates per model, lint/test pass rates, completion time distributions.

  • Works in 6 steps: Run scripts/quality.sh metrics for… → Run scripts/quality.sh pass-rates for… → Run scripts/quality.sh times for… → …
  • The user asks about quality
  • SKILL.md covers When to Use and Instructions
  • Runs Shell scripts from its folder

What it does

Bernstein Quality is an agent skill from sipyourdrink-ltd/bernstein. Show quality metrics for Bernstein runs - success rates per model, lint/test pass rates, completion time distributions. Use when the user asks about quality, reliability, which model performs best, or pass rates.

Its SKILL.md is about 390 tokens, which your agent loads only when the skill is triggered. The skill folder holds 2 other files, including scripts (for example `scripts/quality.sh`).

It sits in Development, covering Linting and formatting. The repository describes itself as: The open‑source AI Agents Governance & Orchestration framework: write the rules declaratively, Bernstein enforces them and produces the verifiable, replayable record. Free…. The licence is Apache-2.0.

When your agent uses it

  • The user asks about quality
  • Which model performs best

Example prompts

  • “/bernstein-quality”

Requirements

  • A Bash shell

Workflow steps

6 steps, taken from the first numbered list in SKILL.md.

  1. Run scripts/quality.sh metrics for overall quality metrics.
  2. Run scripts/quality.sh pass-rates for lint/typecheck/test pass rates by model.
  3. Run scripts/quality.sh times for completion time distributions.
  4. Present a quality dashboard
  5. Highlight any models with significantly lower pass rates.
  6. Recommend model routing adjustments if one model consistently underperforms.

What it can do on your machine

Read from SKILL.md and the folder at commit aee2ea2. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Ships 1 file in scripts/ (Shell), which the agent can run.

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Bernstein Quality loads about 388 tokens when it runs. Until then it costs about 58 tokens; SKILL.md has 102 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~58
When it runs · the whole SKILL.md, loaded when a task matches
~388

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); the scripts in this folder are not scanned.

SKILL.md

The full file from sipyourdrink-ltd/bernstein at commit aee2ea2, republished under its Apache-2.0 licence (© sipyourdrink-ltd). 102 words, ~388 tokens.

Download SKILL.mdSave it as .claude/skills/bernstein-quality/SKILL.md (or your agent's skills folder). This skill also uses 1 other file; get the full folder from GitHub.
name
bernstein-quality
description
Show quality metrics for Bernstein runs - success rates per model, lint/test pass rates, completion time distributions. Use when the user asks about quality, reliability, which model performs best, or pass rates.

Bernstein Quality Metrics

Analyze quality and reliability of agent-generated code.

When to Use

  • User asks "how reliable are the agents?" or "which model is best?"
  • User wants success rates, pass rates, or completion time stats
  • User asks about test failures or lint issues across models
  • User says "show me quality metrics"

Instructions

  1. Run scripts/quality.sh metrics for overall quality metrics.

  2. Run scripts/quality.sh pass-rates for lint/typecheck/test pass rates by model.

  3. Run scripts/quality.sh times for completion time distributions.

  4. Present a quality dashboard:

## Quality Dashboard

### Success Rate by Model
| Model | Tasks | Success | Fail | Rate |
|-------|-------|---------|------|------|
| claude-sonnet-4 | 24 | 22 | 2 | 91.7% |
| gpt-4.1 | 12 | 10 | 2 | 83.3% |

### Pass Rates
| Check | Overall | claude-sonnet-4 | gpt-4.1 |
|-------|---------|-----------------|---------|
| Lint | 96% | 98% | 92% |
| Type-check | 88% | 91% | 83% |
| Tests | 85% | 89% | 75% |

### Completion Times
| Percentile | Time |
|------------|------|
| p50 | 3m 20s |
| p90 | 8m 45s |
| p99 | 15m 12s |
  1. Highlight any models with significantly lower pass rates.
  2. Recommend model routing adjustments if one model consistently underperforms.

© sipyourdrink-ltd, Apache-2.0. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

SKILL.md and 1 other file (scripts) in packages/cursor-plugin/skills/bernstein-quality of sipyourdrink-ltd/bernstein.

  • SKILL.md
  • scripts/quality.sh

Open the folder on GitHubat commit aee2ea2

Compare with similar skills

Bernstein Quality next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Bernstein Quality compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Bernstein Quality this skillsipyourdrink-ltd/bernstein1.4k—~388Automated safety check: PassApache-2.0
Ue Code AuthoringJasonMa0012/MooaToon750—~1.9kAutomated safety check: NotesCustom licence
Procoder Commit Gateazrtydxb/procoder211—~3.7kAutomated safety check: PassApache-2.0
Add Plugin Ruleeslint-config/airbnb-extended130—~646Automated safety check: PassMIT
Lint Repository Markdowncodsen/codsen214—~832Automated safety check: PassMIT
Bmad Builderaj-geddes/claude-code-bmad-skills488—~1.9kAutomated safety check: NotesCustom licence

Similar skills

  • Ue Code Authoring

    JasonMa0012/MooaToon

    A skill your agent uses when writing or modifying UE C++ (classes, actors, components, subsystems, interfaces, function libraries) with Rider MCP available.

    750 GitHub stars~1.9k tokensUpdated 21 days ago
    DevelopmentAuto-check: notes
  • Procoder Commit Gate

    azrtydxb/procoder

    Applies Procoder's senior-developer discipline in a repository: run the commit gate, format through the binary and work through specs, plans and todos.

    211 GitHub stars~3.7k tokensUpdated 11 days ago
    DevelopmentAuto-check passed
  • Add Plugin Rule

    eslint-config/airbnb-extended

    Fix the "<Plugin Updated with <rule" build error from script/checkUpdates.ts by adding a new or deprecated plugin rule to the right rules/ file.

    130 GitHub stars~646 tokensUpdated yesterday
    DevelopmentAuto-check passed
  • Keep the repository's linted Markdown passing npm run lint:markdown.

    214 GitHub stars~832 tokensUpdated 10 days ago
    DevelopmentAuto-check passed
  • Bmad Builder

    aj-geddes/claude-code-bmad-skills

    Meta-skill for scaffolding and validating custom PLANNING/ORCHESTRATION skills within the BMAD Planning & Orchestrator plugin.

    488 GitHub stars~1.9k tokensUpdated 3 mo ago
    DevelopmentAuto-check: notes
  • Coral Debug

    Human-Agent-Society/CORAL

    Verify and debug changes to CORAL itself — smallest reproduce loop per area (grader / daemon / CLI / hooks / manager / workspace / hub / template / config / web), where to look when something breaks…

    1.1k GitHub stars~1.8k tokensUpdated 1 mo ago
    DevelopmentAuto-check passed

More from sipyourdrink-ltd/bernstein

All 9 skills in this repo
  • Bernstein Plan

    sipyourdrink-ltd/bernstein

    Create and manage multi-step execution plans in Bernstein. An agent skill from sipyourdrink-ltd/bernstein.

    1.4k GitHub stars~649 tokensUpdated yesterday
    Auto-check passed
  • Bernstein Run

    sipyourdrink-ltd/bernstein

    Run a verified multi-agent goal with Bernstein. An agent skill from sipyourdrink-ltd/bernstein.

    1.4k GitHub stars~722 tokensUpdated yesterday
    Auto-check passed
  • Bernstein Agents

    sipyourdrink-ltd/bernstein

    Manage Bernstein agents - list active agents, inspect their output, kill stalled agents, or stream live logs.

    1.4k GitHub stars~408 tokensUpdated yesterday
    Auto-check passed
  • Bernstein Alerts

    sipyourdrink-ltd/bernstein

    Show active alerts from Bernstein - failed tasks, stalled agents, budget warnings, blocked tasks needing human intervention.

    1.4k GitHub stars~334 tokensUpdated yesterday
    Auto-check passed
  • Bernstein Approve

    sipyourdrink-ltd/bernstein

    Review and approve/reject pending tasks or plans in Bernstein.

    1.4k GitHub stars~410 tokensUpdated yesterday
    Auto-check passed
  • Bernstein Cost

    sipyourdrink-ltd/bernstein

    Show detailed cost breakdown and budget status for the Bernstein orchestrator.

    1.4k GitHub stars~332 tokensUpdated yesterday
    Auto-check passed

Questions about Bernstein Quality

What does Bernstein Quality do?

Show quality metrics for Bernstein runs - success rates per model, lint/test pass rates, completion time distributions. Bernstein Quality is an agent skill from sipyourdrink-ltd/bernstein. Show quality metrics for Bernstein runs - success rates per model, lint/test pass rates, completion time distributions.

When should I use Bernstein Quality?

Bernstein Quality fits situations like: the user asks about quality; which model performs best.

How do I install Bernstein Quality in Claude Code?

Run `npx skills add sipyourdrink-ltd/bernstein --skill bernstein-quality -a claude-code`. Or copy the skill folder (packages/cursor-plugin/skills/bernstein-quality in sipyourdrink-ltd/bernstein) into .claude/skills/bernstein-quality in your project. Claude Code loads it when a task matches its description.

How do I install Bernstein Quality in Codex?

Run `npx skills add sipyourdrink-ltd/bernstein --skill bernstein-quality -a codex`. Or copy the skill folder (packages/cursor-plugin/skills/bernstein-quality in sipyourdrink-ltd/bernstein) into .agents/skills/bernstein-quality in your project. Codex loads it when a task matches its description.

Can I use Bernstein Quality in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add sipyourdrink-ltd/bernstein --skill bernstein-quality -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/bernstein-quality, .gemini/skills/bernstein-quality, .github/skills/bernstein-quality and .opencode/skills/bernstein-quality in your project.

What does Bernstein Quality need to run?

Going by SKILL.md and its folder, Bernstein Quality needs a shell for the scripts in its folder. Our summary lists: A Bash shell.

Does Bernstein Quality access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is Bernstein Quality safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. The check reads SKILL.md only: the scripts in the folder are not scanned, so read them before running anything.

What licence does Bernstein Quality use?

Bernstein Quality is published under the Apache-2.0 licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Bernstein Quality use?

About 388 tokens (SKILL.md is roughly 1.6k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Bernstein Quality?

Skills that share tags, products or a category with Bernstein Quality: Ue Code Authoring (JasonMa0012/MooaToon, 750 stars), Procoder Commit Gate (azrtydxb/procoder, 211 stars), Add Plugin Rule (eslint-config/airbnb-extended, 130 stars) and Lint Repository Markdown (codsen/codsen, 214 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Bernstein Quality?

sipyourdrink-ltd (a GitHub organization) maintains it in sipyourdrink-ltd/bernstein, which has 1,447 GitHub stars. The repository holds 9 skills in this directory. The repository was last updated on October 9, 2026.

Source: sipyourdrink-ltd/bernstein on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.