Agent skill

Woz Benchmark

by WithWoz in WithWoz/wozcode-plugin

Compare WOZCODE vs vanilla Claude Code on the user's codebase — real cost, turn, and time savings.

No licenceAuto-check: notesDevelopment

Install Woz Benchmark

skills CLI
$ npx skills add WithWoz/wozcode-plugin --skill woz-benchmark -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install WithWoz/wozcode-plugin woz-benchmark --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/WithWoz/wozcode-plugin.git skills-src && mkdir -p .claude/skills && cp -r skills-src/codex/wozcode/skills/woz-benchmark .claude/skills/woz-benchmark && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
woz-benchmark
GitHub stars
209
Token cost
~789 tokens
SKILL.md length
365 words
Files
1
Skills in repo
8
Repo updated
First seen
Licence
None found

At a glance

Compare WOZCODE vs vanilla Claude Code on the user's codebase — real cost, turn, and time savings.

  • Works in 5 steps: Gather inputs — BE BRIEF → Validate the target → Write a temporary benchmark config → …
  • How much does woz save
  • SKILL.md covers Prerequisites and Steps
  • Calls git and node

What it does

Woz Benchmark is an agent skill from WithWoz/wozcode-plugin. Compare WOZCODE vs vanilla Claude Code on the user's codebase — real cost, turn, and time savings. TRIGGER on "compare woz", "how much does woz save", "benchmark woz", "woz vs claude", "show me savings", or /woz-benchmark.

Its SKILL.md is about 790 tokens, which your agent loads only when the skill is triggered. It is a single SKILL.md file with no bundled scripts.

It sits in Development. It works with Git. The repository describes itself as: WOZCODE plugin for Claude Code.

When your agent uses it

  • How much does woz save
  • Show me savings

Example prompts

  • “s codebase — real cost, turn, and time savings. TRIGGER on”
  • “how much does woz save”
  • “benchmark woz”
  • “/woz-benchmark”

Requirements

  • Pre-approved tools (allowed-tools): Bash(node *), Bash(git *), Bash(ls *), Bash(test *), Bash(mkdir *), Bash(date *), Write, Read

Workflow steps

5 steps, taken from the step headings in SKILL.md.

  1. Gather inputs — BE BRIEF
  2. Validate the target
  3. Write a temporary benchmark config
  4. Run the benchmark
  5. Present the results as a savings report

What it can do on your machine

Read from SKILL.md and the folder at commit 1c76873. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves these tools, so the agent can use them without asking each time:

    • Bash(node *)
    • Bash(git *)
    • Bash(ls *)
    • Bash(test *)
    • Bash(mkdir *)
    • Bash(date *)
    • Write
    • Read

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Shell commands in SKILL.md call:

    • git
    • node

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md. Its commands use git, which can reach the network depending on how they are called.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Woz Benchmark loads about 789 tokens when it runs. Until then it costs about 59 tokens; SKILL.md has 365 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~59
When it runs · the whole SKILL.md, loaded when a task matches
~789

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check: notes

The automated check noted patterns worth knowing about, such as sudo or a known installer.

  • NoteMentions a .env fileSKILL.md:26
    eeded, services running, credentials in `.env`)? Skip if the repo is self-contained."

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

Without a licence we can't republish the file, so here is its outline and opening line. It has 365 words (~789 tokens).

“Run a side-by-side comparison of WOZCODE vs vanilla Claude Code on the user's own codebase. Each prompt runs twice against a fresh copy of the repo with git reset --hard between runs, so the target MUST be a clean git…”

— opening of SKILL.md by WithWoz
name
woz-benchmark
allowed-tools
Bash(node *), Bash(git *), Bash(ls *), Bash(test *), Bash(mkdir *), Bash(date *), Write, Read

Read the full SKILL.md on GitHub

Files

Just SKILL.md in codex/wozcode/skills/woz-benchmark of WithWoz/wozcode-plugin.

Open the folder on GitHubat commit 1c76873

Compare with similar skills

Woz Benchmark next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Woz Benchmark compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Woz Benchmark this skillWithWoz/wozcode-plugin209—~789Automated safety check: NotesNone
Finishing a Development Branchobra/superpowers297k5 repos~1.9kAutomated safety check: PassMIT
Code Review ChecklistshareAI-lab/learn-claude-code78k5 repos~1.1kAutomated safety check: PassMIT
Code Design Rationale Investigatorcursor/plugins10k9 repos~2.6kAutomated safety check: PassNone
Contributor-First PR MergeHKUDS/OpenHarness16k1 repos~847Automated safety check: PassMIT
Finishing A Development Branchfarm-fe/farm5.6k34 repos~1.8kAutomated safety check: PassMIT

Similar skills

  • Walks the last step of a branch: confirm tests pass, detect the git environment, ask how to integrate, carry out your choice and clean up the worktree.

    297k GitHub starsUsed in 5 repos~1.9k tokens
    DevelopmentAuto-check passed
  • Code Review Checklist

    shareAI-lab/learn-claude-code

    Reviews code against a five-part checklist covering security, correctness, performance, maintainability and testing, and reports findings in a fixed format.

    78k GitHub starsUsed in 5 repos~1.1k tokens
    DevelopmentAuto-check passed
  • Official

    Digs into why code is shaped the way it is by checking git history, pull requests and connected tools in parallel, then reporting a cited read on the tradeoffs.

    10k GitHub starsUsed in 9 repos~2.6k tokens
    DevelopmentAuto-check passed
  • Merges external GitHub pull requests while keeping the original author credited, and fixes conflicts after the merge instead of rewriting the contribution.

    16k GitHub starsUsed in 1 repo~847 tokens
    DevelopmentAuto-check passed
  • A skill your agent uses when implementation is complete, all tests pass, and you need to decide how to integrate the work - guides completion of development work by presenting structured options for…

    5.6k GitHub starsUsed in 34 repos~1.8k tokens
    DevelopmentAuto-check passed
  • Moves a package from another TryGhost repository into Ghost as an internal workspace package while keeping its Git history, with checkpoints for the steps that need an administrator.

    56k GitHub stars~3.8k tokensUpdated today
    DevelopmentAuto-check passed

More from WithWoz/wozcode-plugin

All 8 skills in this repo
  • Woz Login

    WithWoz/wozcode-plugin

    Authenticate with the Woz service. An agent skill from WithWoz/wozcode-plugin.

    209 GitHub stars~390 tokensUpdated 1 mo ago
    Auto-check passed
  • Woz

    WithWoz/wozcode-plugin

    WOZCODE utilities. An agent skill from WithWoz/wozcode-plugin.

    209 GitHub stars~5.9k tokensUpdated 1 mo ago
    Auto-check: notes
  • Woz Feedback

    WithWoz/wozcode-plugin

    Send feedback or a bug report to the WOZCODE team. An agent skill from WithWoz/wozcode-plugin.

    209 GitHub stars~585 tokensUpdated 1 mo ago
    Auto-check passed
  • Woz Recall

    WithWoz/wozcode-plugin

    Semantically search past Claude Code sessions to recall commands, solutions, and context from prior conversations.

    209 GitHub stars~203 tokensUpdated 1 mo ago
    Auto-check passed
  • Woz Recall

    WithWoz/wozcode-plugin

    Search past Claude Code sessions to recall commands, solutions, and context from prior conversations.

    209 GitHub stars~164 tokensUpdated 1 mo ago
    Auto-check passed
  • Woz Savings

    WithWoz/wozcode-plugin

    Show the WOZCODE savings report — calls saved, time saved, tokens saved, and lifetime totals.

    209 GitHub stars~302 tokensUpdated 1 mo ago
    Auto-check passed

Works with

Categories

Questions about Woz Benchmark

What does Woz Benchmark do?

Compare WOZCODE vs vanilla Claude Code on the user's codebase — real cost, turn, and time savings. Woz Benchmark is an agent skill from WithWoz/wozcode-plugin. Compare WOZCODE vs vanilla Claude Code on the user's codebase — real cost, turn, and time savings.

When should I use Woz Benchmark?

Woz Benchmark fits situations like: how much does woz save; show me savings.

How do I install Woz Benchmark in Claude Code?

Run `npx skills add WithWoz/wozcode-plugin --skill woz-benchmark -a claude-code`. Or copy the skill folder (codex/wozcode/skills/woz-benchmark in WithWoz/wozcode-plugin) into .claude/skills/woz-benchmark in your project. Claude Code loads it when a task matches its description.

How do I install Woz Benchmark in Codex?

Run `npx skills add WithWoz/wozcode-plugin --skill woz-benchmark -a codex`. Or copy the skill folder (codex/wozcode/skills/woz-benchmark in WithWoz/wozcode-plugin) into .agents/skills/woz-benchmark in your project. Codex loads it when a task matches its description.

Can I use Woz Benchmark in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add WithWoz/wozcode-plugin --skill woz-benchmark -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/woz-benchmark, .gemini/skills/woz-benchmark, .github/skills/woz-benchmark and .opencode/skills/woz-benchmark in your project.

What does Woz Benchmark need to run?

Going by SKILL.md and its folder, Woz Benchmark needs the command-line tools its instructions call (git and node). Its frontmatter pre-approves these tools: Bash(node *), Bash(git *), Bash(ls *), Bash(test *), Bash(mkdir *), Bash(date *), Write, Read.

Does Woz Benchmark access the network?

SKILL.md contains no URLs. Its commands use git, which can reach the network depending on how they are called. This is read from the text; nothing was executed.

Is Woz Benchmark safe to install?

Our automated static check of SKILL.md found notes only (mentions a .env file), nothing it rates as a warning. It is not a guarantee. Review the folder before installing.

What licence does Woz Benchmark use?

No licence was found for Woz Benchmark or its repository. Without one, default copyright applies: ask the author before reusing or redistributing it.

How many tokens does Woz Benchmark use?

About 789 tokens (SKILL.md is roughly 3.2k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Woz Benchmark?

Skills that share tags, products or a category with Woz Benchmark: Finishing a Development Branch (obra/superpowers, 297k stars), Code Review Checklist (shareAI-lab/learn-claude-code, 78k stars), Code Design Rationale Investigator (cursor/plugins, 10k stars) and Contributor-First PR Merge (HKUDS/OpenHarness, 16k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Woz Benchmark?

WithWoz (a GitHub organization) maintains it in WithWoz/wozcode-plugin, which has 209 GitHub stars. The repository holds 8 skills in this directory. The repository was last updated on August 17, 2026.

Source: WithWoz/wozcode-plugin on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.