Agent skill

Benchmark

by Log2n-io in Log2n-io/Typhon

Run regression benchmarks, track results, and generate trend reports

Custom licenceAuto-check passed

Install Benchmark

skills CLI
$ npx skills add Log2n-io/Typhon --skill benchmark -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install Log2n-io/Typhon benchmark --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/Log2n-io/Typhon.git skills-src && mkdir -p .claude/skills && cp -r skills-src/.claude/skills/benchmark .claude/skills/benchmark && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
benchmark
GitHub stars
250
Token cost
~1.9k tokens
SKILL.md length
753 words
Files
1
Skills in repo
17
Repo updated
First seen
Licence
Custom licence

At a glance

Run regression benchmarks, track results, and generate trend reports

  • Works in 6 steps: Build in Release → Clean Stale BDN Artifacts → Run Regression Benchmarks → …
  • SKILL.md covers Input and Workflow
  • Calls dotnet and python3

What it does

Benchmark is an agent skill from Log2n-io/Typhon. Run regression benchmarks, track results, and generate trend reports

Its SKILL.md is about 1.9k tokens, which your agent loads only when the skill is triggered. It is a single SKILL.md file with no bundled scripts.

The repository describes itself as: A microsecond-latency ACID database engine with a native Entity-Component-System data model, built for real-time systems.

Example prompts

  • “/benchmark”

Requirements

  • Python 3

Workflow steps

6 steps, taken from the step headings in SKILL.md.

  1. Build in Release
  2. Clean Stale BDN Artifacts
  3. Run Regression Benchmarks
  4. Generate Report
  5. Display Summary
  6. Do NOT publish

What it can do on your machine

Read from SKILL.md and the folder at commit fec10f5. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Shell commands in SKILL.md call:

    • dotnet
    • python3

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Benchmark loads about 1.9k tokens when it runs. Until then it costs about 20 tokens; SKILL.md has 753 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~20
When it runs · the whole SKILL.md, loaded when a task matches
~1.9k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

Its licence (Custom licence) doesn't allow us to republish the file, so here is its outline and opening line. It has 753 words (~1,895 tokens).

“Run Typhon regression benchmarks, record results to history, and generate trend reports with regression detection.”

— opening of SKILL.md by Log2n-io, Custom licence
name
benchmark
argument-hint
[--quick] [--report-only] [--list] [--btree-fast] [--btree-medium] [--btree-full]

Read the full SKILL.md on GitHub

Files

Just SKILL.md in .claude/skills/benchmark of Log2n-io/Typhon.

Open the folder on GitHubat commit fec10f5

Compare with similar skills

Benchmark next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Benchmark compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Benchmark this skillLog2n-io/Typhon250—~1.9kAutomated safety check: PassCustom licence
Benchmarkaffaan-m/ECC274k3 repos~654Automated safety check: PassMIT
Benchmarkaffaan-m/ECC274k—~412Automated safety check: PassMIT
Benchmarkaffaan-m/ECC274k—~330Automated safety check: PassMIT
Visual Regressionthedaviddias/Front-End-Checklist74k—~493Automated safety check: PassMIT
Gstack Performance Benchmarkgarrytan/gstack136k—~7.2kAutomated safety check: NotesMIT

Similar skills

  • Benchmark

    affaan-m/ECC

    Measure performance baselines and detect regressions across browser Core Web Vitals (LCP, INP, CLS, page weight), API endpoint latency percentiles, and build/test feedback times, with before/after…

    274k GitHub starsUsed in 3 repos~654 tokens
    Frontend & DesignAuto-check passed
  • Benchmark

    affaan-m/ECC

    このスキルを使用して、パフォーマンスベースラインを測定し、PR前後の回帰を検出し、スタック代替案を比較します. An agent skill from affaan-m/ECC.

    274k GitHub stars~412 tokensUpdated 2 days ago
    Auto-check passed
  • Benchmark

    affaan-m/ECC

    使用此技能测量性能基线,检测PR前后的回归,并比较堆栈替代方案。

    274k GitHub stars~330 tokensUpdated 2 days ago
    Auto-check passed
  • Visual Regression

    thedaviddias/Front-End-Checklist

    A skill your agent uses when reviewing CI coverage, automated checks, or test strategy related to Use visual regression testing.

    74k GitHub stars~493 tokensUpdated yesterday
    Testing & QAAuto-check passed
  • Establishes page load, Core Web Vitals and resource-size baselines, then compares before and after on every pull request to track performance trends over time.

    136k GitHub stars~7.2k tokensUpdated today
    Frontend & DesignAuto-check: notes
  • Benchmark

    androidx/androidx

    Benchmarking and improving the performance of Jetpack Compose.

    6.1k GitHub stars~1.1k tokensUpdated today
    MobileAuto-check passed

More from Log2n-io/Typhon

All 17 skills in this repo
  • Complete Subtask

    Log2n-io/Typhon

    Complete a sub-issue of an umbrella issue - close it, check parent checkbox, update design doc

    250 GitHub stars~1.5k tokensUpdated yesterday
    Auto-check passed
  • Complete Task

    Log2n-io/Typhon

    Complete work on a GitHub issue - close issue, update artifacts, prompt for doc updates

    250 GitHub stars~2.1k tokensUpdated yesterday
    Auto-check passed
  • Coverage

    Log2n-io/Typhon

    Run code coverage analysis, track class-level results, and generate trend reports

    250 GitHub stars~787 tokensUpdated yesterday
    Auto-check passed
  • Create Issue

    Log2n-io/Typhon

    Create a GitHub issue and add it to the Typhon org project. An agent skill from Log2n-io/Typhon.

    250 GitHub stars~2.6k tokensUpdated yesterday
    Auto-check passed
  • Dev Status

    Log2n-io/Typhon

    Show current development status from GitHub Project. An agent skill from Log2n-io/Typhon.

    250 GitHub stars~874 tokensUpdated yesterday
    Auto-check passed
  • Implement Feature

    Log2n-io/Typhon

    Implement a GitHub issue end-to-end — scope it (whole issue or specific phases), build an acceptance-criteria plan from its design doc, get the plan approved, then develop autonomously with tests…

    250 GitHub stars~2k tokensUpdated yesterday
    Auto-check passed

Questions about Benchmark

What does Benchmark do?

Run regression benchmarks, track results, and generate trend reports. Benchmark is an agent skill from Log2n-io/Typhon.

How do I install Benchmark in Claude Code?

Run `npx skills add Log2n-io/Typhon --skill benchmark -a claude-code`. Or copy the skill folder (.claude/skills/benchmark in Log2n-io/Typhon) into .claude/skills/benchmark in your project. Claude Code loads it when a task matches its description.

How do I install Benchmark in Codex?

Run `npx skills add Log2n-io/Typhon --skill benchmark -a codex`. Or copy the skill folder (.claude/skills/benchmark in Log2n-io/Typhon) into .agents/skills/benchmark in your project. Codex loads it when a task matches its description.

Can I use Benchmark in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add Log2n-io/Typhon --skill benchmark -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/benchmark, .gemini/skills/benchmark, .github/skills/benchmark and .opencode/skills/benchmark in your project.

What does Benchmark need to run?

Going by SKILL.md and its folder, Benchmark needs the command-line tools its instructions call (dotnet and python3). Our summary lists: Python 3.

Does Benchmark access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is Benchmark safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Benchmark use?

Benchmark has a licence file (the repository's licence) that doesn't match a standard licence. Read it on GitHub before reusing the skill.

How many tokens does Benchmark use?

About 1.9k tokens (SKILL.md is roughly 7.6k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Benchmark?

Skills that share tags, products or a category with Benchmark: Benchmark (affaan-m/ECC, 274k stars), Benchmark (affaan-m/ECC, 274k stars), Benchmark (affaan-m/ECC, 274k stars) and Visual Regression (thedaviddias/Front-End-Checklist, 74k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Benchmark?

Log2n-io (a GitHub organization) maintains it in Log2n-io/Typhon, which has 250 GitHub stars. The repository holds 17 skills in this directory. The repository was last updated on October 6, 2026.

Source: Log2n-io/Typhon on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.