Agent skill

Agent Harness Design

by AnastasiyaW in AnastasiyaW/codex-claude-code-config

Designing agent harnesses and tool systems — risk taxonomy for tools, permission decisions, draft/commit pattern, structured tool results, agent budgets (10 types), context trust labels against…

MITAuto-check passedAI & LLM Engineering

Install Agent Harness Design

skills CLI
$ npx skills add AnastasiyaW/codex-claude-code-config --skill agent-harness-design -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install AnastasiyaW/codex-claude-code-config agent-harness-design --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/AnastasiyaW/codex-claude-code-config.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/agent-harness-design .claude/skills/agent-harness-design && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
agent-harness-design
GitHub stars
154
Token cost
~764 tokens
SKILL.md length
216 words
Files
12 (incl. references)
Skills in repo
50
Repo updated
First seen
Licence
MIT

At a glance

Designing agent harnesses and tool systems — risk taxonomy for tools, permission decisions, draft/commit pattern, structured tool results, agent budgets (10 types), context trust labels against…

  • Building a new Agent SDK app
  • Instructions only: no scripts, shell commands, URLs or credentials in SKILL.md
  • Custom orchestrator
  • Cloudflare Worker with tool calls

What it does

Agent Harness Design is an agent skill from AnastasiyaW/codex-claude-code-config. Designing agent harnesses and tool systems — risk taxonomy for tools, permission decisions, draft/commit pattern, structured tool results, agent budgets (10 types), context trust labels against prompt injection, plan-artifact, approval records, observability and traces, evals (13 categories), event model, streaming buffering, 3rd-party skill install checklist, agentic RAG, self-improving SOP loops, model policy, reasoning effort, and Programmatic Tool Calling adoption gates. Use when building a new Agent SDK app…

Its SKILL.md is about 760 tokens, which your agent loads only when the skill is triggered. The skill folder holds 12 other files, including reference files (for example `references/agent-approval-records.md`, `references/agent-budgets.md` and `references/agent-evals.md`).

It sits in AI & LLM Engineering, covering Autonomous loops, LLM evaluation and Structured output and tool calling. It works with Model Context Protocol and Cloudflare Workers. The repository describes itself as: Claude Code, Codex, and multi-agent configuration system: principles, hooks, skills, and workflow patterns for AI-assisted development. The licence is MIT.

When your agent uses it

  • Building a new Agent SDK app
  • Custom orchestrator
  • Cloudflare Worker with tool calls
  • Agentic RAG pipeline

Example prompts

  • “/agent-harness-design”

What it can do on your machine

Read from SKILL.md and the folder at commit 67709af. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    No scripts in the folder and no shell commands in SKILL.md.

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Agent Harness Design loads about 764 tokens when it runs, and up to ~26k if it reads all its reference files. Until then it costs about 241 tokens; SKILL.md has 216 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~241
When it runs · the whole SKILL.md, loaded when a task matches
~764
With references · SKILL.md plus every file in references/, read only if the agent opens them
~26k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from AnastasiyaW/codex-claude-code-config at commit 67709af, republished under its MIT licence (© AnastasiyaW). 216 words, ~764 tokens.

Download SKILL.mdSave it as .claude/skills/agent-harness-design/SKILL.md (or your agent's skills folder). This skill also uses 11 other files; get the full folder from GitHub.
name
agent-harness-design
description
Designing agent harnesses and tool systems — risk taxonomy for tools, permission decisions, draft/commit pattern, structured tool results, agent budgets (10 types), context trust labels against prompt injection, plan-artifact, approval records, observability and traces, evals (13 categories), event model, streaming buffering, 3rd-party skill install checklist, agentic RAG, self-improving SOP loops, model policy, reasoning effort, and Programmatic Tool Calling adoption gates. Use when building a new Agent SDK app, custom orchestrator, MCP server, Cloudflare Worker with tool calls, agentic RAG pipeline, model router, or model-tier policy; when designing tools and permissions; when writing an agent loop; or when you need trust labels for external content. Do NOT use for improving or auditing an already-built harness (use harness-audit / harness-design instead), nor for ordinary Claude Code sessions where the harness is already given.

Agent Harness Design

Eleven operational reference sheets for designing a safe, observable agent harness. They are situational — load only the one(s) relevant to the current task from references/ (this is why they live in a skill rather than always-on rules: building an agent harness is occasional, so the detail should not bloat every session's context).

  • references/agent-tool-design.md — 15-class risk taxonomy, 7-type permission decision object, draft/commit naming, structured tool results, deferred tool loading, hosted vs client tools, connector code-execution pattern.
  • references/context-trust-labels.md — trusted / semi_trusted / untrusted labels + verbatim boundary statement; prompt-injection defense.
  • references/agent-budgets.md — 10 mandatory budget types every agent loop must declare.
  • references/agent-evals.md — 13 eval categories + 13 adversarial test cases + when to add regression evals.
  • references/agent-observability.md — 16 trace fields per model call, 7-question audit, 6-step incident response.
  • references/agentic-rag-model-policy.md — self-improving agentic RAG state, specialist roles, evaluation vectors, Pareto selection, OpenAI model/effort policy, and Programmatic Tool Calling adoption gates.
  • references/agent-plan-artifact.md — planning mode, plan artifact format (10 fields), plan-validate-execute.
  • references/agent-approval-records.md — approval request/result JSON schemas, scope/expiration, no self-approval.
  • references/agent-streaming.md — buffering for incremental tool calls when stream=True; abort handling; output guardrail modes.
  • references/agent-event-model.md — 13 typed events for harness state persistence (replay/audit/compaction/evals).
  • references/agent-skill-install-checklist.md — pre/during/post install + audit + incident response for 3rd-party skills.

Source: distilled from the agents-best-practices skill (Denis Sergeevitch, MIT) + Anthropic harness-design engineering. Read the specific reference before applying — do not work from this index alone.

© AnastasiyaW, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

SKILL.md and 11 other files (references) in skills/agent-harness-design of AnastasiyaW/codex-claude-code-config.

  • SKILL.md
  • references/agent-approval-records.md
  • references/agent-budgets.md
  • references/agent-evals.md
  • references/agent-event-model.md
  • references/agent-observability.md
  • references/agent-plan-artifact.md
  • references/agent-skill-install-checklist.md
  • references/agent-streaming.md
  • references/agent-tool-design.md
  • references/agentic-rag-model-policy.md
  • references/context-trust-labels.md

Open the folder on GitHubat commit 67709af

Compare with similar skills

Agent Harness Design next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Agent Harness Design compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Agent Harness Design this skillAnastasiyaW/codex-claude-code-config154—~764Automated safety check: PassMIT
Building Agentsericrisco/rsc-harness180—~5kAutomated safety check: PassMIT
Building Agent Systemstelagod/code-abyss244—~691Automated safety check: PassMIT
LangchainOrchestra-Research/AI-Research-SKILLs13k2 repos~3.2kAutomated safety check: PassMIT
Agentsop HTTP Tool Wrappingagentsope/SkillAlchemy436—~5.9kAutomated safety check: PassMIT
Agent Evalericrisco/rsc-harness190—~3.2kAutomated safety check: PassMIT

Similar skills

  • Building Agents

    ericrisco/rsc-harness

    A skill your agent uses when building or restructuring an LLM agent — provider adapter, tool calling, structured output, RAG, agent loop, eval gate, cost routing, tracing, MCP server —…

    180 GitHub stars~5k tokensUpdated yesterday
    AI & LLM EngineeringAuto-check passed
  • Building Agent Systems

    telagod/code-abyss

    AI agent and LLM system engineering reference covering single-agent dev (ReAct, tool calling, plan-execute), multi-agent coordination (swarm, role decomposition, file locking), LLM security (prompt…

    244 GitHub stars~691 tokensUpdated 2 mo ago
    AI & LLM EngineeringAuto-check passed
  • Langchain

    Orchestra-Research/AI-Research-SKILLs

    Framework for building LLM-powered applications with agents, chains, and RAG.

    13k GitHub starsUsed in 2 repos~3.2k tokens
    AI & LLM EngineeringAuto-check passed
  • Agentsop HTTP Tool Wrapping

    agentsope/SkillAlchemy

    Decision protocol for wrapping a REST / GraphQL / RPC API as a tool an LLM agent can call.

    436 GitHub stars~5.9k tokensUpdated 2 days ago
    AI & LLM EngineeringAuto-check passed
  • Agent Eval

    ericrisco/rsc-harness

    A skill your agent uses when measuring whether an LLM or agent system actually got better and gating merges on it: golden sets, fixing an inflated LLM-as-judge, scoring RAG (faithfulness, contextual…

    190 GitHub stars~3.2k tokensUpdated today
    AI & LLM EngineeringAuto-check passed
  • Agent Squad Python Guide

    2FastLabs/agent-squad

    Map of the agent-squad Python framework for async multi-agent orchestration: which agent, classifier, storage and tool provider to pick, and the pitfalls to avoid.

    7.8k GitHub stars~4.7k tokensUpdated 3 days ago
    AI & LLM EngineeringAuto-check passed

More from AnastasiyaW/codex-claude-code-config

All 50 skills in this repo
  • Bug Reproducer

    AnastasiyaW/codex-claude-code-config

    Find likely software bugs in a codebase, rank concrete bug candidates, and prove or reject them with focused regression tests before proposing a fix.

    154 GitHub stars~4.1k tokensUpdated today
    Auto-check passed
  • Motion Framer

    AnastasiyaW/codex-claude-code-config

    A skill your agent uses when implementing Motion or Framer Motion in React/JavaScript: interactive UI components, micro-interactions, gestures, layout or page transitions, and scroll-based animation.

    154 GitHub starsUsed in 1 repo~5.2k tokens
    Auto-check passed
  • Proof Verify

    AnastasiyaW/codex-claude-code-config

    Plan-based verification - freeze acceptance criteria before building, then verify after with an independent fresh-context agent (the builder must not verify their own work).

    154 GitHub stars~2.6k tokensUpdated today
    Auto-check passed
  • Workflow Orchestration

    AnastasiyaW/codex-claude-code-config

    Написание и запуск Claude Code dynamic workflows (JS-оркестратор субагентов).

    154 GitHub stars~3.8k tokensUpdated today
    Auto-check passed
  • Notebooklm Grounded Research

    AnastasiyaW/codex-claude-code-config

    A skill your agent uses when: NotebookLM, notebooklm MCP, large documentation sets, courses, books, papers, or citation-backed research are mentioned.

    154 GitHub stars~2.4k tokensUpdated today
    Auto-check: warnings
  • Deepseek Provider Contract

    AnastasiyaW/codex-claude-code-config

    Validate a proposed DeepSeek API integration before any key or project context is sent: check thinking-mode tool-call history, strict-schema assumptions, bounded output, and provider data boundaries.

    154 GitHub stars~1.2k tokensUpdated today
    Auto-check passed

Questions about Agent Harness Design

What does Agent Harness Design do?

Designing agent harnesses and tool systems — risk taxonomy for tools, permission decisions, draft/commit pattern, structured tool results, agent budgets (10 types), context trust labels against…. Agent Harness Design is an agent skill from AnastasiyaW/codex-claude-code-config. Designing agent harnesses and tool systems — risk taxonomy for tools, permission decisions, draft/commit pattern, structured tool results, agent budgets (10 types), context trust labels against prompt injection, plan-artifact, approval records, observability and traces, evals (13 categories), event model, streaming buffering, 3rd-party skill install checklist, agentic RAG, self-improving SOP loops, model policy, reasoning effort, and Programmatic Tool Calling adoption gates.

When should I use Agent Harness Design?

Agent Harness Design fits situations like: building a new Agent SDK app; custom orchestrator; cloudflare Worker with tool calls; agentic RAG pipeline.

How do I install Agent Harness Design in Claude Code?

Run `npx skills add AnastasiyaW/codex-claude-code-config --skill agent-harness-design -a claude-code`. Or copy the skill folder (skills/agent-harness-design in AnastasiyaW/codex-claude-code-config) into .claude/skills/agent-harness-design in your project. Claude Code loads it when a task matches its description.

How do I install Agent Harness Design in Codex?

Run `npx skills add AnastasiyaW/codex-claude-code-config --skill agent-harness-design -a codex`. Or copy the skill folder (skills/agent-harness-design in AnastasiyaW/codex-claude-code-config) into .agents/skills/agent-harness-design in your project. Codex loads it when a task matches its description.

Can I use Agent Harness Design in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add AnastasiyaW/codex-claude-code-config --skill agent-harness-design -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/agent-harness-design, .gemini/skills/agent-harness-design, .github/skills/agent-harness-design and .opencode/skills/agent-harness-design in your project.

What does Agent Harness Design need to run?

SKILL.md names no scripts, command-line tools or credentials: Agent Harness Design is instructions for the agent only.

Does Agent Harness Design access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is Agent Harness Design safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Agent Harness Design use?

Agent Harness Design is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Agent Harness Design use?

About 764 tokens (SKILL.md is roughly 3.1k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full. Its references folder adds about 25k tokens, read only when the agent opens those files.

What are the alternatives to Agent Harness Design?

Skills that share tags, products or a category with Agent Harness Design: Building Agents (ericrisco/rsc-harness, 180 stars), Building Agent Systems (telagod/code-abyss, 244 stars), Langchain (Orchestra-Research/AI-Research-SKILLs, 13k stars) and Agentsop HTTP Tool Wrapping (agentsope/SkillAlchemy, 436 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Agent Harness Design?

AnastasiyaW (a GitHub user) maintains it in AnastasiyaW/codex-claude-code-config, which has 154 GitHub stars. The repository holds 50 skills in this directory. The repository was last updated on October 9, 2026.

Source: AnastasiyaW/codex-claude-code-config on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.