Agent skill

Code Specialist

by benchflow-ai in benchflow-ai/benchflow

Delegate complex coding tasks to a specialist model. An agent skill from benchflow-ai/benchflow.

Apache-2.0Auto-check passedDevelopment

Install Code Specialist

skills CLI
$ npx skills add benchflow-ai/benchflow --skill code-specialist -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install benchflow-ai/benchflow code-specialist --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/benchflow-ai/benchflow.git skills-src && mkdir -p .claude/skills && cp -r skills-src/benchmarks/models-as-skills .claude/skills/code-specialist && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
code-specialist
GitHub stars
353
Token cost
~225 tokens
SKILL.md length
91 words
Files
3
Skills in repo
8
Repo updated
First seen
Licence
Apache-2.0

At a glance

Delegate complex coding tasks to a specialist model. An agent skill from benchflow-ai/benchflow.

  • Works in 3 steps: What the code should do (input → output) → Constraints (time/space complexity,… → Edge cases to handle
  • Facing algorithmic challenges
  • SKILL.md covers When to delegate and How to use
  • Instructions only: no scripts, shell commands, URLs or credentials in SKILL.md

What it does

Code Specialist is an agent skill from benchflow-ai/benchflow. Delegate complex coding tasks to a specialist model. Use when facing algorithmic challenges, performance optimization, or tricky debugging that benefits from focused code expertise.

Its SKILL.md is about 230 tokens, which your agent loads only when the skill is triggered. The skill folder holds 3 other files (for example `evals/evals.json` and `models-as-skills.md`).

It sits in Development, covering Performance optimization. The repository describes itself as: Research infra for creating RL environments, post-training, and evals. The licence is Apache-2.0.

When your agent uses it

  • Facing algorithmic challenges
  • Performance optimization
  • Tricky debugging that benefits from focused code expertise

Example prompts

  • “/code-specialist”

Workflow steps

3 steps, taken from the first numbered list in SKILL.md.

  1. What the code should do (input → output)
  2. Constraints (time/space complexity, language)
  3. Edge cases to handle

What it can do on your machine

Read from SKILL.md and the folder at commit e965eee. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    No scripts in the folder and no shell commands in SKILL.md.

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Code Specialist loads about 225 tokens when it runs. Until then it costs about 49 tokens; SKILL.md has 91 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~49
When it runs · the whole SKILL.md, loaded when a task matches
~225

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from benchflow-ai/benchflow at commit e965eee, republished under its Apache-2.0 licence (© benchflow-ai). 91 words, ~225 tokens.

Download SKILL.mdSave it as .claude/skills/code-specialist/SKILL.md (or your agent's skills folder). This skill also uses 2 other files; get the full folder from GitHub.
name
code-specialist
description
Delegate complex coding tasks to a specialist model. Use when facing algorithmic challenges, performance optimization, or tricky debugging that benefits from focused code expertise.

Code Specialist

When facing a complex coding task, delegate to the specialist rather than solving it yourself.

When to delegate

  • Algorithm implementation requiring specific knowledge (graph algorithms, dynamic programming)
  • Performance optimization (O(n²) → O(n log n) conversions)
  • Debugging complex race conditions or memory issues
  • Code generation requiring precise syntax (regex, SQL, shell)

How to use

Describe the problem clearly to the specialist. Include:

  1. What the code should do (input → output)
  2. Constraints (time/space complexity, language)
  3. Edge cases to handle

The specialist returns code that you should integrate into your solution.

© benchflow-ai, Apache-2.0. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

SKILL.md and 2 other files in benchmarks/models-as-skills of benchflow-ai/benchflow.

  • SKILL.md
  • evals/evals.json
  • models-as-skills.md

Open the folder on GitHubat commit e965eee

Compare with similar skills

Code Specialist next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Code Specialist compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Code Specialist this skillbenchflow-ai/benchflow353—~225Automated safety check: PassApache-2.0
Code Review ChecklistshareAI-lab/learn-claude-code78k5 repos~1.1kAutomated safety check: PassMIT
LLM Torch Profiler Analysissgl-project/sglang37k2 repos~6.4kAutomated safety check: PassApache-2.0
Pycrazyguitar/pysheeet8.2k—~886Automated safety check: PassMIT
Cmux Debugging Guidemanaflow-ai/cmux28k1 repos~1.1kAutomated safety check: PassCustom licence
Electron Heap Snapshot Analysiskeybase/client9.3k—~875Automated safety check: PassBSD-3-Clause

Similar skills

  • Code Review Checklist

    shareAI-lab/learn-claude-code

    Reviews code against a five-part checklist covering security, correctness, performance, maintainability and testing, and reports findings in a fixed format.

    78k GitHub starsUsed in 5 repos~1.1k tokens
    DevelopmentAuto-check passed
  • LLM Torch Profiler Analysis

    sgl-project/sglang

    Unified LLM torch-profiler triage skill for sglang, vllm, TensorRT-LLM, and TokenSpeed.

    37k GitHub starsUsed in 2 repos~6.4k tokens
    DevelopmentAuto-check passed
  • Py

    crazyguitar/pysheeet

    Comprehensive Python programming reference covering syntax, concurrency, networking, databases, ML/LLM development, and HPC.

    8.2k GitHub stars~886 tokensUpdated yesterday
    DevelopmentAuto-check passed
  • Cmux Debugging Guide

    manaflow-ai/cmux

    Covers debug logging, the Debug menu, profiling rules and runtime pitfalls for working on the cmux macOS terminal app.

    28k GitHub starsUsed in 1 repo~1.1k tokens
    DevelopmentAuto-check passed
  • Analyzes V8, Chrome and Electron .heapsnapshot files with Node scripts to find memory leaks, detached DOM nodes and the retainer paths that keep objects alive.

    9.3k GitHub stars~875 tokensUpdated today
    DevelopmentAuto-check passed
  • Runs controlled JMH experiments on the Caffeine cache to find shared contention and hot-path waste, then reviews correctness and returns a reviewable patch.

    18k GitHub stars~2.6k tokensUpdated 2 days ago
    DevelopmentAuto-check: notes

More from benchflow-ai/benchflow

All 8 skills in this repo
  • Task Creator

    benchflow-ai/benchflow

    SkillsBench task authoring — walk a contributor from idea to submission-ready task following CONTRIBUTING.md and the task-implementation rubric.

    353 GitHub stars~4.5k tokensUpdated yesterday
    Auto-check passed
  • Task Review

    benchflow-ai/benchflow

    SkillsBench task PR review — classifies the task track (standard / research / multimodal), runs static policy checks against the track-specific rubric, benchmarks the task across oracle plus Claude…

    353 GitHub stars~4.5k tokensUpdated yesterday
    Auto-check: notes
  • Benchflow Experiment Review

    benchflow-ai/benchflow

    Review Benchflow or SkillsBench task-run trajectories and integration-test Benchflow code changes.

    353 GitHub stars~4k tokensUpdated yesterday
    Auto-check passed
  • Benchflow

    benchflow-ai/benchflow

    Run agent benchmarks, create tasks, analyze results, and manage agents using BenchFlow.

    353 GitHub stars~1.9k tokensUpdated yesterday
    Auto-check: notes
  • Benchflow Traj Upload Ops

    benchflow-ai/benchflow

    Operate, test, troubleshoot, and explain bench traj upload for public or trusted-direct trajectory contributions, including interactive and fully specified commands, dry runs, input validation…

    353 GitHub stars~1.1k tokensUpdated yesterday
    Auto-check passed
  • Benchflow Traj Upload

    benchflow-ai/benchflow

    Find a local Claude Code or Codex session, open the BenchFlow trajectory viewer, and submit it after the user reviews it.

    353 GitHub stars~1.9k tokensUpdated yesterday
    Auto-check: notes

Categories

Questions about Code Specialist

What does Code Specialist do?

Delegate complex coding tasks to a specialist model. An agent skill from benchflow-ai/benchflow. Code Specialist is an agent skill from benchflow-ai/benchflow. Delegate complex coding tasks to a specialist model.

When should I use Code Specialist?

Code Specialist fits situations like: facing algorithmic challenges; performance optimization; tricky debugging that benefits from focused code expertise.

How do I install Code Specialist in Claude Code?

Run `npx skills add benchflow-ai/benchflow --skill code-specialist -a claude-code`. Or copy the skill folder (benchmarks/models-as-skills in benchflow-ai/benchflow) into .claude/skills/code-specialist in your project. Claude Code loads it when a task matches its description.

How do I install Code Specialist in Codex?

Run `npx skills add benchflow-ai/benchflow --skill code-specialist -a codex`. Or copy the skill folder (benchmarks/models-as-skills in benchflow-ai/benchflow) into .agents/skills/code-specialist in your project. Codex loads it when a task matches its description.

Can I use Code Specialist in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add benchflow-ai/benchflow --skill code-specialist -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/code-specialist, .gemini/skills/code-specialist, .github/skills/code-specialist and .opencode/skills/code-specialist in your project.

What does Code Specialist need to run?

SKILL.md names no scripts, command-line tools or credentials: Code Specialist is instructions for the agent only.

Does Code Specialist access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is Code Specialist safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Code Specialist use?

Code Specialist is published under the Apache-2.0 licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Code Specialist use?

About 225 tokens (SKILL.md is roughly 900 characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Code Specialist?

Skills that share tags, products or a category with Code Specialist: Code Review Checklist (shareAI-lab/learn-claude-code, 78k stars), LLM Torch Profiler Analysis (sgl-project/sglang, 37k stars), Py (crazyguitar/pysheeet, 8.2k stars) and Cmux Debugging Guide (manaflow-ai/cmux, 28k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Code Specialist?

benchflow-ai (a GitHub organization) maintains it in benchflow-ai/benchflow, which has 353 GitHub stars. The repository holds 8 skills in this directory. The repository was last updated on October 6, 2026.

Source: benchflow-ai/benchflow on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.