Agent skill

Benchmark

by samchon in samchon/typia

Defines typia benchmark fixture integrity, result reporting, and publication safeguards.

MITAuto-check passedAI & LLM Engineering

Install Benchmark

skills CLI
$ npx skills add samchon/typia --skill benchmark -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install samchon/typia benchmark --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/samchon/typia.git skills-src && mkdir -p .claude/skills && cp -r skills-src/.agents/skills/benchmark .claude/skills/benchmark && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
benchmark
GitHub stars
5.9k
Token cost
~1.2k tokens
SKILL.md length
592 words
Files
2
Skills in repo
6
Repo updated
First seen
Licence
MIT

At a glance

Defines typia benchmark fixture integrity, result reporting, and publication safeguards.

  • Works in 3 steps: Run the relevant template entrypoint,… → Confirm every comparator receives the… → Review the generated-program diff. A…
  • AI & LLM Engineering work in your project
  • SKILL.md covers Measurement Integrity, Fixture Changes, Report Results and Campaign Batching, plus 1 more section
  • Calls git and pnpm

What it does

Benchmark is an agent skill from samchon/typia. Defines typia benchmark fixture integrity, result reporting, and publication safeguards. Use before running or modifying @typia/benchmark, changing a fixture, or publishing benchmark results.

Its SKILL.md is about 1.2k tokens, which your agent loads only when the skill is triggered. The skill folder holds 1 other file (for example `performance.md`).

It sits in AI & LLM Engineering. The repository describes itself as: Super-fast/easy runtime validators and serializers via transformation. The licence is MIT.

When your agent uses it

  • AI & LLM Engineering work in your project

Example prompts

  • “Use the benchmark skill to define typia benchmark fixture integrity, result reporting, and publication safeguards”
  • “/benchmark”

Workflow steps

3 steps, taken from the first numbered list in SKILL.md.

  1. Run the relevant template entrypoint, then the benchmark package's build commands until they finish green. pnpm template regenerates only…
  2. Confirm every comparator receives the same structure and input for the measured row.
  3. Review the generated-program diff. A stale or inconsistent generated program contaminates every later run.

What it can do on your machine

Read from SKILL.md and the folder at commit d9d3436. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Shell commands in SKILL.md call:

    • git
    • pnpm

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md. Its commands use git and pnpm, which can reach the network depending on how they are called.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Benchmark loads about 1.2k tokens when it runs. Until then it costs about 50 tokens; SKILL.md has 592 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~50
When it runs · the whole SKILL.md, loaded when a task matches
~1.2k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from samchon/typia at commit d9d3436, republished under its MIT licence (© samchon). 592 words, ~1,178 tokens.

Download SKILL.mdSave it as .claude/skills/benchmark/SKILL.md (or your agent's skills folder). This skill also uses 1 other file; get the full folder from GitHub.
name
benchmark
description
Defines typia benchmark fixture integrity, result reporting, and publication safeguards. Use before running or modifying @typia/benchmark, changing a fixture, or publishing benchmark results.

Benchmark

This repository owns one benchmark system. Read its procedure in full before acting:

  • performance.md: @typia/benchmark throughput against competing runtime libraries, including generated programs, structure fixtures, and the per-CPU result archive.

Measurement Integrity

  • Measure the real product. Do not add benchmark-only branches, fixture-name checks, expected-answer checks, monkey patches, or agent restrictions that would be wrong for an unmeasured repository.
  • Give every comparator the setup its own documentation prescribes. Measuring a deliberately underconfigured competitor invalidates the comparison.
  • Preserve the workload defined by the selected procedure. A faster result obtained by validating, serializing, serving, or reading less input is not an optimization.
  • Treat a surprising result as evidence that the change is not yet understood. Inspect the raw report or trace before accepting, explaining away, or patching around it.

Fixture Changes

benchmark/src/programs/ mixes generated leaf programs with hand-maintained helpers. Edit generated benchmark-*.ts files through their source under benchmark/src/template/; edit shared create*.ts helpers and benchmark/src/structures/ in place.

Finish every fixture change before publishing it:

  1. Run the relevant template entrypoint, then the benchmark package's build commands until they finish green. pnpm template regenerates only the templates imported by benchmark/src/template/index.ts. Defer repository formatting to the development skill's final pre-merge cleanup gate.
  2. Confirm every comparator receives the same structure and input for the measured row.
  3. Review the generated-program diff. A stale or inconsistent generated program contaminates every later run.

Benchmark READMEs and prose follow AGENTS.md ## Maintenance and the documentation skill.

Report Results

Every result table reported in chat or committed to the result archive must be preserved for the active pull request. When the user has authorized PR updates under the pull-request skill, maintain one sticky comment beginning with <!-- typia-benchmark-results -->; update it with the latest table, report paths, and known invalid or missing categories.

If no pull request exists or no update is authorized, keep the result in the final report and mark the comment as pending. Post it only after the user creates or authorizes updating the pull request.

Show full SKILL.md (266 more words)Show less

Campaign Batching

When benchmark findings produce multiple implementation issues, load the issue-campaign skill and use its planning and claim procedure unchanged for the dependency DAG, claim freeze, and official GitHub createdAt-to-mergedAt duration: Plan One Cycle Pull Request for a solo campaign, or Plan And Claim A Pull Request Wave with its batch admission test, merge pressure, and grouping and split ledger when the user explicitly requested a parallel campaign. This benchmark skill continues to own workload, fixture, measurement, result, and publication integrity; do not redefine pull-request batching here.

Campaign Cleanup

When a benchmark campaign uses a disposable worktree or an isolated measurement root, finish cleanup before marking that assignment complete. Preserve the committed result archive and compact command evidence, but remove every disposable worktree and its assigned mutable roots: GOCACHE, GOTMPDIR, TTSC_CACHE_DIR, generated-program scratch tree, dependency install tree, and temporary consumer or report staging tree. Go temporary assets are never reusable campaign evidence.

  1. Record the measured commit, command, result paths and hashes, environment, and any retained published result archive.
  2. Confirm the worktree and mutable roots contain no unreported source or result artifact.
  3. Remove the mutable roots and, for an assigned disposable worktree, run git worktree remove --force <path> for its exact path.
  4. Verify every removed root and, when applicable, worktree path no longer exists, run git worktree prune, delete the associated disposable local topic branch when one was created, and confirm no worktree registration remains.

If a measurement is abandoned or invalid, retain only the diagnostic record needed to explain it; remove its worktree and Go temporary assets by the same procedure.

© samchon, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

SKILL.md and 1 other file in .agents/skills/benchmark of samchon/typia.

  • SKILL.md
  • performance.md

Open the folder on GitHubat commit d9d3436

Compare with similar skills

Benchmark next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Benchmark compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Benchmark this skillsamchon/typia5.9k—~1.2kAutomated safety check: PassMIT
Doc AuthorInsForge/InsForge13k—~3.7kAutomated safety check: PassMIT
Add AI Chat Toolryokun6/ryos1.3k—~2.2kAutomated safety check: PassAGPL-3.0
Agent Interface DesignNeeeophytee/finding-unknowns-skills343—~650Automated safety check: PassMIT
Search Benchmarkagentic-community/mcp-gateway-registry962—~1.9kAutomated safety check: PassApache-2.0
Implement Featureljxpython/ai-agent-platform137—~1kAutomated safety check: PassNone

Similar skills

  • Doc Author

    InsForge/InsForge

    Write, edit, and maintain documentation. An agent skill from InsForge/InsForge.

    13k GitHub stars~3.7k tokensUpdated yesterday
    AI & LLM EngineeringAuto-check passed
  • Add AI Chat Tool

    ryokun6/ryos

    Add or modify an AI chat tool ("Ask Ryo" capability) in ryOS.

    1.3k GitHub stars~2.2k tokensUpdated today
    AI & LLM EngineeringAuto-check passed
  • Agent Interface Design

    Neeeophytee/finding-unknowns-skills

    Design tools, scripts, and CLIs that an agent will call, so the interface teaches its own use instead of a wall of prose and examples.

    343 GitHub stars~650 tokensUpdated 9 days ago
    AI & LLM EngineeringAuto-check passed
  • Search Benchmark

    agentic-community/mcp-gateway-registry

    Generate a search quality benchmark for the AI Registry. An agent skill from agentic-community/mcp-gateway-registry.

    962 GitHub stars~1.9k tokensUpdated yesterday
    AI & LLM EngineeringAuto-check passed
  • Implement Feature

    ljxpython/ai-agent-platform

    AI 在项目文档(docs/projects/{YYYYMMDD}-{项目名}/)已存在时,实现任务过程中自动调用(不需要用户手动触发),记录改动细节到 implementation/ 目录。跳过条件:单项目改动的小改动不需要创建实现记录。

    137 GitHub stars~1k tokensUpdated yesterday
    AI & LLM EngineeringAuto-check passed
  • Autogpt Agents

    Orchestra-Research/AI-Research-SKILLs

    Autonomous AI agent platform for building and deploying continuous agents.

    13k GitHub starsUsed in 3 repos~2.3k tokens
    AI & LLM EngineeringAuto-check: notes

More from samchon/typia

  • Project

    samchon/typia

    Defines the typia product contract, workspace layout, package boundaries, and canonical commands.

    5.9k GitHub stars~1.8k tokensUpdated yesterday
    Auto-check passed
  • Development

    samchon/typia

    Defines typia implementation rules, testing standards, validation, consequence analysis, and change integrity.

    5.9k GitHub stars~4.9k tokensUpdated yesterday
    Auto-check: notes
  • Issue Campaign

    samchon/typia

    Defines the default solo repository-wide issue campaign for typia: exhaustive discovery, lead-vetted issue publication, one unified CI-validated implementation pull request per cycle, solo…

    5.9k GitHub stars~3.1k tokensUpdated yesterday
    Auto-check passed
  • Review

    samchon/typia

    Defines exhaustive solo review, Self-Review, and solo repository-wide issue-discovery rounds for typia.

    5.9k GitHub stars~2.5k tokensUpdated yesterday
    Auto-check passed
  • Documentation

    samchon/typia

    Defines README, website-guide, and agent-instruction structure, audience, prose formatting, and voice for typia.

    5.9k GitHub stars~1.1k tokensUpdated yesterday
    Auto-check passed

Questions about Benchmark

What does Benchmark do?

Defines typia benchmark fixture integrity, result reporting, and publication safeguards. Benchmark is an agent skill from samchon/typia. Defines typia benchmark fixture integrity, result reporting, and publication safeguards.

When should I use Benchmark?

Benchmark fits situations like: AI & LLM Engineering work in your project.

How do I install Benchmark in Claude Code?

Run `npx skills add samchon/typia --skill benchmark -a claude-code`. Or copy the skill folder (.agents/skills/benchmark in samchon/typia) into .claude/skills/benchmark in your project. Claude Code loads it when a task matches its description.

How do I install Benchmark in Codex?

Run `npx skills add samchon/typia --skill benchmark -a codex`. Or copy the skill folder (.agents/skills/benchmark in samchon/typia) into .agents/skills/benchmark in your project. Codex loads it when a task matches its description.

Can I use Benchmark in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add samchon/typia --skill benchmark -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/benchmark, .gemini/skills/benchmark, .github/skills/benchmark and .opencode/skills/benchmark in your project.

What does Benchmark need to run?

Going by SKILL.md and its folder, Benchmark needs the command-line tools its instructions call (git and pnpm).

Does Benchmark access the network?

SKILL.md contains no URLs. Its commands use git, which can reach the network depending on how they are called. This is read from the text; nothing was executed.

Is Benchmark safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Benchmark use?

Benchmark is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Benchmark use?

About 1.2k tokens (SKILL.md is roughly 4.7k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Benchmark?

Skills that share tags, products or a category with Benchmark: Doc Author (InsForge/InsForge, 13k stars), Add AI Chat Tool (ryokun6/ryos, 1.3k stars), Agent Interface Design (Neeeophytee/finding-unknowns-skills, 343 stars) and Search Benchmark (agentic-community/mcp-gateway-registry, 962 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Benchmark?

samchon (a GitHub user) maintains it in samchon/typia, which has 5,935 GitHub stars. The repository holds 6 skills in this directory. The repository was last updated on October 6, 2026.

Source: samchon/typia on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.