Agent skill

Agent Design Review

by mohitagw15856 in mohitagw15856/pm-claude-skills

Review an LLM agent design and find where it will be unreliable, expensive, or unsafe.

MITAuto-check passedMedia & Creative

Install Agent Design Review

skills CLI
$ npx skills add mohitagw15856/pm-claude-skills --skill agent-design-review -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install mohitagw15856/pm-claude-skills agent-design-review --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/mohitagw15856/pm-claude-skills.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/agent-design-review .claude/skills/agent-design-review && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
agent-design-review
GitHub stars
1.4k
Token cost
~1.2k tokens
SKILL.md length
560 words
Files
1
Skills in repo
1,348
Repo updated
First seen
Licence
MIT

At a glance

Review an LLM agent design and find where it will be unreliable, expensive, or unsafe.

  • Asked to review an agent architecture
  • SKILL.md covers Working from a brief, Required Inputs, Output Format and Quality Checks, plus 3 more sections
  • Instructions only: no scripts, shell commands, URLs or credentials in SKILL.md
  • Critique a multi-step/tool-using agent

What it does

Agent Design Review is an agent skill from mohitagw15856/pm-claude-skills. Review an LLM agent design and find where it will be unreliable, expensive, or unsafe. Use when asked to review an agent architecture, critique a multi-step/tool-using agent, debug an agent that loops or goes off-task, or harden an agent before launch. Produces a structured review — task fit, control flow, tools, memory/context, failure handling, cost, and safety — with prioritised findings and fixes.

Its SKILL.md is about 1.2k tokens, which your agent loads only when the skill is triggered. It is a single SKILL.md file with no bundled scripts.

It sits in Media & Creative, covering Design review and critique. The repository describes itself as: 1255 professional Agent Skills for Claude, ChatGPT, Gemini, Cursor & Codex — PRDs, postmortems, leases, medical bills, layoffs, go-bags, new countries. Plain markdown, MIT, in… The licence is MIT.

When your agent uses it

  • Asked to review an agent architecture
  • Critique a multi-step/tool-using agent
  • Debug an agent that loops
  • Harden an agent before launch

Example prompts

  • “/agent-design-review”

What it can do on your machine

Read from SKILL.md and the folder at commit 1cbf1f0. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    No scripts in the folder and no shell commands in SKILL.md.

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Agent Design Review loads about 1.2k tokens when it runs. Until then it costs about 106 tokens; SKILL.md has 560 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~106
When it runs · the whole SKILL.md, loaded when a task matches
~1.2k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from mohitagw15856/pm-claude-skills at commit 1cbf1f0, republished under its MIT licence (© mohitagw15856). 560 words, ~1,155 tokens.

Download SKILL.mdSave it as .claude/skills/agent-design-review/SKILL.md (or your agent's skills folder).
name
agent-design-review
description
Review an LLM agent design and find where it will be unreliable, expensive, or unsafe. Use when asked to review an agent architecture, critique a multi-step/tool-using agent, debug an agent that loops or goes off-task, or harden an agent before launch. Produces a structured review — task fit, control flow, tools, memory/context, failure handling, cost, and safety — with prioritised findings and fixes.

Agent Design Review Skill

Most agents don't fail because the model is weak — they fail because the design lets them loop, call the wrong tool, lose the thread across steps, or burn tokens with no stopping rule. This skill reviews an agent's architecture against the decisions that actually determine reliability, and ranks the fixes — so "it works in the demo but not in prod" becomes a specific list of changes. (Writing a new agent spec? Use agent-spec.)

Working from a brief

Given a sketch ("a research agent that searches, reads, and writes a report"), deliver the full review anyway — infer the likely control flow and tools, label the inference, and flag what to confirm. Never withhold the review for missing detail.

Required Inputs

Ask for these only if they aren't already provided (else infer and label):

  • What the agent does — its goal, and what a successful run produces.
  • Control flow — single prompt, plan-then-execute, ReAct loop, or multi-agent; and the stopping condition.
  • Tools & actions — what it can call, and which actions have side effects (write, send, pay).
  • Memory & context — what state carries across steps, and how context is kept in budget.
  • Constraints — latency, cost per run, and the trust boundary (untrusted input? real-world actions?).

Output Format

Agent Review: [agent]

1. Summary — will this be reliable in production? The top 3 risks and the single change that helps most.

2. Findings by dimension — for each, what's sound and what's fragile:

DimensionFindingSeverityFix
Control flowno max-steps / no progress check → loopsHighstep budget + "am I making progress?" check + halt
Tool useoverlapping tools confuse selectionMedfewer, sharply-described tools; allowlist
Contextfull history re-sent each step → cost + driftHighsummarise/scope memory per step
Failure handlingone tool error aborts the runMedretry/backoff + graceful degradation
Safetyacts without confirmation on writesHighhuman/confirm gate on side-effecting actions

3. Reliability checklist — termination guarantee (it always stops), error recovery, idempotency of side-effecting actions, and determinism where it matters.

4. Cost & latency — where tokens/steps are spent and how to cut them (cheaper model for sub-steps, caching, fewer round-trips) without losing quality. Pair with llm-cost-latency-budget.

5. Safety — untrusted input/tool output handled as data not instructions, least-privilege tools, and confirmation gates on high-impact actions. Pair with llm-guardrails-spec.

6. Prioritised fix plan — ordered by impact-to-effort.

Show full SKILL.md (181 more words)Show less

Quality Checks

  • The agent has a guaranteed stopping condition (step/budget cap + progress check) — no unbounded loops
  • Side-effecting actions are idempotent or gated by a confirmation
  • Tools are few and sharply described so selection is unambiguous; access is least-privilege
  • Context strategy keeps the window in budget across steps (no naive full-history resend)
  • Tool errors are recovered, not fatal — retry/backoff and graceful degradation
  • Findings are severity-ranked and the fix plan is ordered by impact

Anti-Patterns

  • Do not approve an agent with no termination guarantee — "it usually stops" is an outage waiting to happen
  • Do not let it take irreversible actions without a confirmation gate
  • Do not give it many overlapping tools — selection accuracy drops as the toolset grows
  • Do not resend the whole history every step — cost and drift both climb
  • Do not treat tool/retrieved output as trusted instructions — it's the injection surface

Based On

LLM agent design practice — bounded control flow, least-privilege tool use, context management, error recovery, and safety gating.

Example Trigger Phrases

  • "Review an agent architecture."
  • "Critique a multi-step/tool-using agent."
  • "Debug an agent that loops."
  • "Harden an agent before launch."

© mohitagw15856, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

Just SKILL.md in skills/agent-design-review of mohitagw15856/pm-claude-skills.

Open the folder on GitHubat commit 1cbf1f0

Compare with similar skills

Agent Design Review next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Agent Design Review compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Agent Design Review this skillmohitagw15856/pm-claude-skills1.4k—~1.2kAutomated safety check: PassMIT
Consult ClaudeEpicenterHQ/epicenter4.8k—~2kAutomated safety check: PassCustom licence
System Atlasinkboard/system-atlas430—~2.3kAutomated safety check: PassMIT
Design Image Studiokangarooking/design-image-studio102—~1.5kAutomated safety check: PassMIT
Kicad Reviewmixelpixx/Konnect927—~3.2kAutomated safety check: PassAGPL-3.0
Design AuditUniClipboard/UniClipboard1.9k—~554Automated safety check: PassAGPL-3.0

Similar skills

  • Consult Claude

    EpicenterHQ/epicenter

    Assign Claude Code a read-only investigation, recommendation, or finished text draft.

    4.8k GitHub stars~2k tokensUpdated 3 days ago
    Media & CreativeAuto-check passed
  • System Atlas

    inkboard/system-atlas

    Build and maintain an explorable, progressively-disclosed isometric "atlas" of a system's architecture — an interactive page (hover to read, click to pin, go inside for steps, moving data packets…

    430 GitHub stars~2.3k tokensUpdated 1 mo ago
    Media & CreativeAuto-check passed
  • Design Image Studio

    kangarooking/design-image-studio

    Directly generate design-oriented AI images with strong creative direction and prompt engineering.

    102 GitHub stars~1.5k tokensUpdated 5 mo ago
    Media & CreativeAuto-check passed
  • Kicad Review

    mixelpixx/Konnect

    Design review and validation workflow for KiCAD projects via MCP tools.

    927 GitHub stars~3.2k tokensUpdated 5 days ago
    Media & CreativeAuto-check passed
  • Design Audit

    UniClipboard/UniClipboard

    定期审计代码库的工程设计问题(高心智复杂度、单一真相源被破坏、catch-all 胖接口、死代码、散落魔法字面量、泄漏抽象、资源生命周期靠环形缓冲)与可优化点,范围限定为自上次审计以来的 git churn,每条发现都落到 file:line 并对照本项目自己的 VISION.md / 各级 AGENTS.md / memory…

    1.9k GitHub stars~554 tokensUpdated today
    Media & CreativeAuto-check passed
  • L1 AI Design Review

    PaperMoonuu/Design-workflow-skills

    L1 × AI 设计评审:对已完成的单页、局部 UI 设计稿进行小型迭代评审,识别影响面、状态遗漏、文案与一致性风险,并给出 P0/P1/P2 建议和验收清单。用户提供 Figma 链接、截图、前后设计稿或可评审原型,并要求设计走查、风险评审或开发前 UI 检查时使用;不用于设计前方案预检、完整多页面流程或 L2 开发交付。

    316 GitHub stars~492 tokensUpdated 19 days ago
    Media & CreativeAuto-check passed

More from mohitagw15856/pm-claude-skills

All 1,348 skills in this repo
  • Car Tco

    mohitagw15856/pm-claude-skills

    Compare the total cost of car ownership across buy-new, buy-used, lease, and keep-your-current-car — depreciation, insurance, maintenance ramp, and fuel over a real horizon, not just the monthly…

    1.4k GitHub stars~1.1k tokensUpdated 2 days ago
    Auto-check passed
  • Cs Health Scorecard

    mohitagw15856/pm-claude-skills

    Build a customer health scorecard for a specific account. An agent skill from mohitagw15856/pm-claude-skills.

    1.4k GitHub stars~2.4k tokensUpdated 2 days ago
    Auto-check passed
  • Exit Waterfall

    mohitagw15856/pm-claude-skills

    Compute who gets what at each exit price from a cap table — liquidation preferences, conversion points, and where the founders' share collapses.

    1.4k GitHub stars~1.1k tokensUpdated 2 days ago
    Auto-check passed
  • Feature Prioritisation

    mohitagw15856/pm-claude-skills

    Apply prioritisation frameworks (RICE, MoSCoW, Kano, ICE, Opportunity Scoring) to rank features and backlog items.

    1.4k GitHub stars~2k tokensUpdated 2 days ago
    Auto-check passed
  • Fire Number

    mohitagw15856/pm-claude-skills

    Compute a financial-independence (FIRE) target and years-to-reach with every assumption labeled as an assumption — plus a sensitivity table instead of a single false-precision answer.

    1.4k GitHub stars~1.1k tokensUpdated 2 days ago
    Auto-check passed
  • Freelance Rate

    mohitagw15856/pm-claude-skills

    Derive a freelance day/hourly rate backwards from target income, honest billable utilization, overhead, and the self-employment tax premium — the arithmetic that proves a rate is not salary÷2000.

    1.4k GitHub stars~1.2k tokensUpdated 2 days ago
    Auto-check passed

Questions about Agent Design Review

What does Agent Design Review do?

Review an LLM agent design and find where it will be unreliable, expensive, or unsafe. Agent Design Review is an agent skill from mohitagw15856/pm-claude-skills. Review an LLM agent design and find where it will be unreliable, expensive, or unsafe.

When should I use Agent Design Review?

Agent Design Review fits situations like: asked to review an agent architecture; critique a multi-step/tool-using agent; debug an agent that loops; harden an agent before launch.

How do I install Agent Design Review in Claude Code?

Run `npx skills add mohitagw15856/pm-claude-skills --skill agent-design-review -a claude-code`. Or copy the skill folder (skills/agent-design-review in mohitagw15856/pm-claude-skills) into .claude/skills/agent-design-review in your project. Claude Code loads it when a task matches its description.

How do I install Agent Design Review in Codex?

Run `npx skills add mohitagw15856/pm-claude-skills --skill agent-design-review -a codex`. Or copy the skill folder (skills/agent-design-review in mohitagw15856/pm-claude-skills) into .agents/skills/agent-design-review in your project. Codex loads it when a task matches its description.

Can I use Agent Design Review in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add mohitagw15856/pm-claude-skills --skill agent-design-review -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/agent-design-review, .gemini/skills/agent-design-review, .github/skills/agent-design-review and .opencode/skills/agent-design-review in your project.

What does Agent Design Review need to run?

SKILL.md names no scripts, command-line tools or credentials: Agent Design Review is instructions for the agent only.

Does Agent Design Review access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is Agent Design Review safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Agent Design Review use?

Agent Design Review is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Agent Design Review use?

About 1.2k tokens (SKILL.md is roughly 4.6k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Agent Design Review?

Skills that share tags, products or a category with Agent Design Review: Consult Claude (EpicenterHQ/epicenter, 4.8k stars), System Atlas (inkboard/system-atlas, 430 stars), Design Image Studio (kangarooking/design-image-studio, 102 stars) and Kicad Review (mixelpixx/Konnect, 927 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Agent Design Review?

mohitagw15856 (a GitHub user) maintains it in mohitagw15856/pm-claude-skills, which has 1,434 GitHub stars. The repository holds 1,348 skills in this directory. The repository was last updated on October 9, 2026.

Source: mohitagw15856/pm-claude-skills on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.