Agent skill

Jmh Benchmark Compare

by eclipse-rdf4j in eclipse-rdf4j/rdf4j

Parse JMH result text by finding the first header line that starts with Benchmark and contains Mode and Score, build a structured table for all columns/rows, compare overlapping benchmarks across 2+…

BSD-3-ClauseAuto-check passedDocuments & Office

Install Jmh Benchmark Compare

skills CLI
$ npx skills add eclipse-rdf4j/rdf4j --skill jmh-benchmark-compare -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install eclipse-rdf4j/rdf4j jmh-benchmark-compare --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/eclipse-rdf4j/rdf4j.git skills-src && mkdir -p .claude/skills && cp -r skills-src/.agent/skills/jmh-benchmark-compare .claude/skills/jmh-benchmark-compare && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
jmh-benchmark-compare
GitHub stars
420
Token cost
~804 tokens
SKILL.md length
230 words
Files
7 (incl. scripts, references)
Skills in repo
6
Repo updated
First seen
Licence
BSD-3-Clause

At a glance

Parse JMH result text by finding the first header line that starts with Benchmark and contains Mode and Score, build a structured table for all columns/rows, compare overlapping benchmarks across 2+…

  • Works in 5 steps: Detect first JMH table header line → Derive column boundaries from that header. → Parse all following benchmark rows into… → …
  • Benchmark run comparisons
  • SKILL.md covers Quick start, Core behavior, Inputs and overlap and Filters and regression shortcuts, plus 3 more sections
  • Runs Python scripts from its folder; calls python3

What it does

Jmh Benchmark Compare is an agent skill from eclipse-rdf4j/rdf4j. Parse JMH result text by finding the first header line that starts with Benchmark and contains Mode and Score, build a structured table for all columns/rows, compare overlapping benchmarks across 2+ files, compute Diff Score and Diff %, filter by deviation or regression thresholds, analyze regressions over time from filename/mtime timestamps, and export sortable reports to txt/md/csv/xlsx/html. Use for benchmark run comparisons, regression triage, and directory-wide historical analysis.

Its SKILL.md is about 800 tokens, which your agent loads only when the skill is triggered. The skill folder holds 9 other files, including scripts and reference files (for example `agents/openai.yaml`, `references/timestamps-and-discovery.md` and `scripts/jmh_benchmark_compare.py`).

It sits in Documents & Office, covering Excel spreadsheets and CSV and tabular files. It works with Microsoft Excel. The repository describes itself as: Eclipse RDF4J: scalable RDF for Java. The licence is BSD-3-Clause.

When your agent uses it

  • Benchmark run comparisons
  • Regression triage
  • Directory-wide historical analysis

Example prompts

  • “/jmh-benchmark-compare”

Requirements

  • Python 3

Workflow steps

5 steps, taken from the first numbered list in SKILL.md.

  1. Detect first JMH table header line
  2. Derive column boundaries from that header.
  3. Parse all following benchmark rows into an internal table.
  4. Match overlapping benchmark keys across files.
  5. Add derived columns

What it can do on your machine

Read from SKILL.md and the folder at commit fdbf6e5. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Ships 4 files in scripts/ (Python), which the agent can run.

    Shell commands in SKILL.md call:

    • python3

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Jmh Benchmark Compare loads about 804 tokens when it runs, and up to ~994 if it reads all its reference files. Until then it costs about 128 tokens; SKILL.md has 230 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~128
When it runs · the whole SKILL.md, loaded when a task matches
~804
With references · SKILL.md plus every file in references/, read only if the agent opens them
~994

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); the scripts in this folder are not scanned.

SKILL.md

The full file from eclipse-rdf4j/rdf4j at commit fdbf6e5, republished under its BSD-3-Clause licence (© eclipse-rdf4j). 230 words, ~804 tokens.

Download SKILL.mdSave it as .claude/skills/jmh-benchmark-compare/SKILL.md (or your agent's skills folder). This skill also uses 6 other files; get the full folder from GitHub.
name
jmh-benchmark-compare
description
Parse JMH result text by finding the first header line that starts with Benchmark and contains Mode and Score, build a structured table for all columns/rows, compare overlapping benchmarks across 2+ files, compute Diff Score and Diff %, filter by deviation or regression thresholds, analyze regressions over time from filename/mtime timestamps, and export sortable reports to txt/md/csv/xlsx/html. Use for benchmark run comparisons, regression triage, and directory-wide historical analysis.

jmh-benchmark-compare

Use this skill when benchmark output comparison must be reproducible, sortable, and exportable.

Quick start

Run two-file comparison:

bash
python3 .codex/skills/jmh-benchmark-compare/scripts/jmh_benchmark_compare.py \
  /path/run-a.txt /path/run-b.txt \
  --export-formats txt,md,csv,xlsx,html \
  --output-dir /tmp \
  --output-base jmh-compare

Sort by diff percent (descending):

bash
python3 .codex/skills/jmh-benchmark-compare/scripts/jmh_benchmark_compare.py \
  run-a.txt run-b.txt \
  --sort-column "Diff % [run-b - run-a]" \
  --sort-desc \
  --export-formats md \
  --output /tmp/jmh-diff.md

Core behavior

  1. Detect first JMH table header line: line.startswith("Benchmark") and "Mode" in line and "Score" in line.
  2. Derive column boundaries from that header.
  3. Parse all following benchmark rows into an internal table.
  4. Match overlapping benchmark keys across files.
  5. Add derived columns: Diff Score [target - baseline], Diff % [target - baseline], Status [...].

Default key columns are all columns except Cnt, Score, Error. Override via --id-columns.

Inputs and overlap

  • Pass any mix of files and directories.
  • Directory entries are scanned for files that contain a JMH header.
  • --overlap-mode all keeps only rows present in all files.
  • --overlap-mode any keeps rows present in at least two files.
  • Baseline selection: --baseline <index-or-label>.

Filters and regression shortcuts

  • Hide tiny deltas: --min-deviation-pct 1.0
  • Show only regressions above threshold: --regressions-over-pct 3.0
  • Control direction interpretation: --score-direction auto|higher|lower

Historical analysis

Analyze trends across many runs:

bash
python3 .codex/skills/jmh-benchmark-compare/scripts/jmh_benchmark_compare.py \
  /path/bench-history \
  --recursive \
  --glob "*.txt" \
  --timestamp-source auto \
  --analyze-over-time \
  --regressions-over-pct 2.5 \
  --export-formats html,csv \
  --output-dir /tmp \
  --output-base jmh-history

Timeline report files are emitted with -timeline suffix.

Exports

  • txt: aligned plain-text table.
  • md: valid markdown table.
  • csv: spreadsheet-friendly CSV.
  • xlsx: native Excel workbook (single sheet). (xslx alias accepted)
  • html: sortable table (click header), built-in CSS + JS, color theme selector.

If one format and explicit destination needed, use --output /path/file.ext. If multiple formats, use --output-dir + --output-base.

Script

scripts/jmh_benchmark_compare.py

For timestamp parsing behavior and filename examples, see: references/timestamps-and-discovery.md

© eclipse-rdf4j, BSD-3-Clause. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

SKILL.md and 6 other files (scripts, references) in .agent/skills/jmh-benchmark-compare of eclipse-rdf4j/rdf4j.

  • SKILL.md
  • agents/openai.yaml
  • references/timestamps-and-discovery.md
  • scripts/jmh_benchmark_compare.py
  • scripts/jmh_compare_core.py
  • scripts/jmh_compare_export.py
  • scripts/test_jmh_compare_core.py

Open the folder on GitHubat commit fdbf6e5

Compare with similar skills

Jmh Benchmark Compare next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Jmh Benchmark Compare compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Jmh Benchmark Compare this skilleclipse-rdf4j/rdf4j420—~804Automated safety check: PassBSD-3-Clause
Data Table Managern8n-io/n8n207k—~2.3kAutomated safety check: PassCustom licence
Instrument Data To Allotropeaws-samples/amazon-bedrock-agents-healthcare-lifesciences2742 repos~2.7kAutomated safety check: PassApache-2.0
Markitshift-labs-ai/markit1.3k—~299Automated safety check: PassMIT
Convert Fileduckdb/duckdb-skills6031 repos~720Automated safety check: NotesMIT
Research Integrity Auditxuzhougeng/wisp-science1k—~2.6kAutomated safety check: PassAGPL-3.0

Similar skills

  • Official

    Load before calling data-tables or parse-file. An agent skill from n8n-io/n8n.

    207k GitHub stars~2.3k tokensUpdated today
    Documents & OfficeAuto-check passed
  • Instrument Data To Allotrope

    aws-samples/amazon-bedrock-agents-healthcare-lifesciences

    Official

    Convert laboratory instrument output files (PDF, CSV, Excel, TXT) to Allotrope Simple Model (ASM) JSON format or flattened 2D CSV.

    274 GitHub starsUsed in 2 repos~2.7k tokens
    Documents & OfficeAuto-check passed
  • Markit

    shift-labs-ai/markit

    Convert files and URLs to Markdown. An agent skill from shift-labs-ai/markit.

    1.3k GitHub stars~299 tokensUpdated 1 mo ago
    Documents & OfficeAuto-check passed
  • Convert File

    duckdb/duckdb-skills

    Official

    Convert any data file to another format: CSV, Parquet, JSON, Excel, GeoJSON, and more.

    603 GitHub starsUsed in 1 repo~720 tokens
    Documents & OfficeAuto-check: notes
  • Research Integrity Audit

    xuzhougeng/wisp-science

    学术审查 / research-integrity screening of a manuscript's figures and reported numbers.

    1k GitHub stars~2.6k tokensUpdated today
    Documents & OfficeAuto-check passed
  • Excel Parser

    Harryoung/efka

    Smart Excel/CSV file parsing with intelligent routing based on file complexity analysis.

    104 GitHub stars~2.3k tokensUpdated 6 mo ago
    Documents & OfficeAuto-check passed

More from eclipse-rdf4j/rdf4j

  • Docker Jfr Benchmark Loop

    eclipse-rdf4j/rdf4j

    Run a repeatable RDF4J performance loop against one JMH benchmark in Docker with Linux Java 26 and JFR CPU-time profiling.

    420 GitHub stars~945 tokensUpdated 2 days ago
    Auto-check passed
  • Mvnf

    eclipse-rdf4j/rdf4j

    Run Maven tests in this repo with a consistent workflow (module test-artifact cleanup, root -Pquick install to refresh .m2repo, then module verify or a single test class/method).

    420 GitHub stars~971 tokensUpdated 2 days ago
    Auto-check passed
  • Query Plan Snapshot CLI

    eclipse-rdf4j/rdf4j

    Use QueryPlanSnapshotCli to capture and compare RDF4J query plans, then assess likely performance improvements/regressions from execution verification and semantic plan diffs.

    420 GitHub stars~1.5k tokensUpdated 2 days ago
    Auto-check passed
  • Debug Surefire

    eclipse-rdf4j/rdf4j

    Debug Maven Surefire unit tests by running them in JDWP "wait for debugger" mode (-Dmaven.surefire.debug) and attaching to the forked test JVM using jdb (preferred for CLI/agent debugging)…

    420 GitHub stars~2.4k tokensUpdated 2 days ago
    Auto-check passed
  • Gh Read Inspector

    eclipse-rdf4j/rdf4j

    Retrieve GitHub issues, pull requests, and milestones with read-only, whitelisted gh commands only.

    420 GitHub stars~950 tokensUpdated 2 days ago
    Auto-check passed

Works with

Questions about Jmh Benchmark Compare

What does Jmh Benchmark Compare do?

Parse JMH result text by finding the first header line that starts with Benchmark and contains Mode and Score, build a structured table for all columns/rows, compare overlapping benchmarks across 2+…. Jmh Benchmark Compare is an agent skill from eclipse-rdf4j/rdf4j. Parse JMH result text by finding the first header line that starts with Benchmark and contains Mode and Score, build a structured table for all columns/rows, compare overlapping benchmarks across 2+ files, compute Diff Score and Diff %, filter by deviation or regression thresholds, analyze regressions over time from filename/mtime timestamps, and export sortable reports to txt/md/csv/xlsx/html.

When should I use Jmh Benchmark Compare?

Jmh Benchmark Compare fits situations like: benchmark run comparisons; regression triage; directory-wide historical analysis.

How do I install Jmh Benchmark Compare in Claude Code?

Run `npx skills add eclipse-rdf4j/rdf4j --skill jmh-benchmark-compare -a claude-code`. Or copy the skill folder (.agent/skills/jmh-benchmark-compare in eclipse-rdf4j/rdf4j) into .claude/skills/jmh-benchmark-compare in your project. Claude Code loads it when a task matches its description.

How do I install Jmh Benchmark Compare in Codex?

Run `npx skills add eclipse-rdf4j/rdf4j --skill jmh-benchmark-compare -a codex`. Or copy the skill folder (.agent/skills/jmh-benchmark-compare in eclipse-rdf4j/rdf4j) into .agents/skills/jmh-benchmark-compare in your project. Codex loads it when a task matches its description.

Can I use Jmh Benchmark Compare in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add eclipse-rdf4j/rdf4j --skill jmh-benchmark-compare -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/jmh-benchmark-compare, .gemini/skills/jmh-benchmark-compare, .github/skills/jmh-benchmark-compare and .opencode/skills/jmh-benchmark-compare in your project.

What does Jmh Benchmark Compare need to run?

Going by SKILL.md and its folder, Jmh Benchmark Compare needs Python for the scripts in its folder and the command-line tools its instructions call (python3). Our summary lists: Python 3.

Does Jmh Benchmark Compare access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is Jmh Benchmark Compare safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. The check reads SKILL.md only: the scripts in the folder are not scanned, so read them before running anything.

What licence does Jmh Benchmark Compare use?

Jmh Benchmark Compare is published under the BSD-3-Clause licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Jmh Benchmark Compare use?

About 804 tokens (SKILL.md is roughly 3.2k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full. Its references folder adds about 190 tokens, read only when the agent opens those files.

What are the alternatives to Jmh Benchmark Compare?

Skills that share tags, products or a category with Jmh Benchmark Compare: Data Table Manager (n8n-io/n8n, 207k stars), Instrument Data To Allotrope (aws-samples/amazon-bedrock-agents-healthcare-lifesciences, 274 stars), Markit (shift-labs-ai/markit, 1.3k stars) and Convert File (duckdb/duckdb-skills, 603 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Jmh Benchmark Compare?

eclipse-rdf4j (a GitHub organization) maintains it in eclipse-rdf4j/rdf4j, which has 420 GitHub stars. The repository holds 6 skills in this directory. The repository was last updated on October 9, 2026.

Source: eclipse-rdf4j/rdf4j on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.