Arena
Jakeschincariol/arena-skill
Make 100 versions of Claude fight to the death over one task.
Plan review criteria — the six review axes, severity rubric, approval standard, and output format for reviewing implementation plans.
$ npx skills add penpot/penpot --skill plan-review-criteria -a claude-codeProject install by default; add -g for ~/.claude/skills/.
$ gh skill install penpot/penpot plan-review-criteria --agent claude-codeProject scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).
$ git clone --depth 1 https://github.com/penpot/penpot.git skills-src && mkdir -p .claude/skills && cp -r skills-src/.agents/skills/plan-review-criteria .claude/skills/plan-review-criteria && rm -rf skills-srcUse ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.
Claude Code skills documentation · loads skills from .claude/skills/
Install the "plan-review-criteria" agent skill from https://github.com/penpot/penpot/tree/develop/.agents/skills/plan-review-criteria into .claude/skills/plan-review-criteria/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "plan-review-criteria", then confirm the skill loads.Claude Code copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$skill-installer install https://github.com/penpot/penpot/tree/develop/.agents/skills/plan-review-criteriaType this inside Codex. $skill-installer <name> installs a curated skill from openai/skills. The installer writes to $CODEX_HOME/skills (default ~/.codex/skills). Restart Codex if the skill does not show up.
$ npx skills add penpot/penpot --skill plan-review-criteria -a codexProject install goes to .agents/skills/; add -g for ~/.codex/skills/.
$ gh skill install penpot/penpot plan-review-criteria --agent codexProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/penpot/penpot.git skills-src && mkdir -p .agents/skills && cp -r skills-src/.agents/skills/plan-review-criteria .agents/skills/plan-review-criteria && rm -rf skills-srcUse ~/.agents/skills/ instead of .agents/skills for a personal install.
Codex skills documentation · loads skills from .agents/skills/
Install the "plan-review-criteria" agent skill from https://github.com/penpot/penpot/tree/develop/.agents/skills/plan-review-criteria into .agents/skills/plan-review-criteria/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "plan-review-criteria", then confirm the skill loads.Codex copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add penpot/penpot --skill plan-review-criteria -a cursorProject install goes to .agents/skills/; add -g for ~/.cursor/skills/.
$ gh skill install penpot/penpot plan-review-criteria --agent cursorProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/penpot/penpot.git skills-src && mkdir -p .cursor/skills && cp -r skills-src/.agents/skills/plan-review-criteria .cursor/skills/plan-review-criteria && rm -rf skills-srcUse ~/.cursor/skills/ instead of .cursor/skills for a personal install.
Cursor skills documentation · loads skills from .cursor/skills/, .agents/skills/, .claude/skills/, .codex/skills/
Install the "plan-review-criteria" agent skill from https://github.com/penpot/penpot/tree/develop/.agents/skills/plan-review-criteria into .cursor/skills/plan-review-criteria/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "plan-review-criteria", then confirm the skill loads.Cursor copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gemini skills install https://github.com/penpot/penpot.git --path .agents/skills/plan-review-criteria--scope user (default) or --scope workspace; --path is the subfolder of the repo that holds the skill; --consent skips the security confirmation prompt.
$ npx skills add penpot/penpot --skill plan-review-criteria -a gemini-cliProject install goes to .agents/skills/; add -g for ~/.gemini/skills/.
$ gh skill install penpot/penpot plan-review-criteria --agent gemini-cliProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/penpot/penpot.git skills-src && mkdir -p .gemini/skills && cp -r skills-src/.agents/skills/plan-review-criteria .gemini/skills/plan-review-criteria && rm -rf skills-srcUse ~/.gemini/skills/ instead of .gemini/skills for a personal install, then run /skills reload.
Gemini CLI skills documentation · loads skills from .gemini/skills/, .agents/skills/
Install the "plan-review-criteria" agent skill from https://github.com/penpot/penpot/tree/develop/.agents/skills/plan-review-criteria into .gemini/skills/plan-review-criteria/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "plan-review-criteria", then confirm the skill loads.Gemini CLI copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gh skill install penpot/penpot plan-review-criteriaInstalls for Copilot at project scope by default; add --scope user for a personal install. Preview a skill first with gh skill preview. Needs GitHub CLI 2.90.0 or later (public preview).
$ npx skills add penpot/penpot --skill plan-review-criteria -a github-copilotProject install goes to .agents/skills/; add -g for ~/.copilot/skills/.
$ git clone --depth 1 https://github.com/penpot/penpot.git skills-src && mkdir -p .github/skills && cp -r skills-src/.agents/skills/plan-review-criteria .github/skills/plan-review-criteria && rm -rf skills-srcUse ~/.copilot/skills/ instead of .github/skills for a personal install. Commit .github/skills so cloud agent and code review can use it.
GitHub Copilot skills documentation · loads skills from .github/skills/, .claude/skills/, .agents/skills/
Install the "plan-review-criteria" agent skill from https://github.com/penpot/penpot/tree/develop/.agents/skills/plan-review-criteria into .github/skills/plan-review-criteria/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "plan-review-criteria", then confirm the skill loads.GitHub Copilot copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add penpot/penpot --skill plan-review-criteria -a opencodeOpenCode documents no install command of its own. Project install goes to .agents/skills/; add -g for ~/.config/opencode/skills/.
$ gh skill install penpot/penpot plan-review-criteria --agent opencodeProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/penpot/penpot.git skills-src && mkdir -p .opencode/skills && cp -r skills-src/.agents/skills/plan-review-criteria .opencode/skills/plan-review-criteria && rm -rf skills-srcUse ~/.config/opencode/skills/ instead of .opencode/skills for a personal install.
OpenCode skills documentation · loads skills from .opencode/skills/, .claude/skills/, .agents/skills/
Install the "plan-review-criteria" agent skill from https://github.com/penpot/penpot/tree/develop/.agents/skills/plan-review-criteria into .opencode/skills/plan-review-criteria/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "plan-review-criteria", then confirm the skill loads.OpenCode copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
plan-review-criteriaPlan review criteria — the six review axes, severity rubric, approval standard, and output format for reviewing implementation plans.
Plan Review Criteria is an agent skill from penpot/penpot. Plan review criteria — the six review axes, severity rubric, approval standard, and output format for reviewing implementation plans. Loaded by the reviewer subagent of the review-plan flow. Not a user-facing flow — to review a plan, use the review-plan flow.
Its SKILL.md is about 3.3k tokens, which your agent loads only when the skill is triggered. It is a single SKILL.md file with no bundled scripts.
It sits in Agent Workflows, covering Accessibility, Planning and Quizzes and assessments. The repository describes itself as: Penpot: The open-source design platform for Product teams that need scalable collaboration. The licence is MPL-2.0.
12 steps, taken from the step headings in SKILL.md.
Read from SKILL.md and the folder at commit 10955f1. It shows what the files ask for, not the result of running them.
Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.
From allowed-tools in the SKILL.md frontmatter.
No scripts in the folder and no shell commands in SKILL.md (its code samples are markdown).
From the folder's file list and the shell code blocks in SKILL.md.
No URLs in SKILL.md.
From URLs in SKILL.md, links to its own repository left out.
Names no API keys, tokens, secrets or passwords.
From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.
Plan Review Criteria loads about 3.3k tokens when it runs. Until then it costs about 70 tokens; SKILL.md has 1,296 words of instructions outside code blocks.
Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.
The automated check found no risky patterns in SKILL.md.
Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.
The full file from penpot/penpot at commit 10955f1, republished under its MPL-2.0 licence (© penpot). 1,296 words, ~3,278 tokens.
.claude/skills/plan-review-criteria/SKILL.md (or your agent's skills folder).Multi-dimensional plan review with quality gates. Every plan gets reviewed before implementation starts — no exceptions. Review covers six axes: completeness, task quality, architecture & sequencing, risk coverage, actionability, and proposed code quality.
The approval standard: Approve a plan when it is specific enough that a skilled implementer could execute it without guessing, the task ordering is sound, and risks are acknowledged. Perfect plans don't exist — the goal is confidence that implementation won't derail. Don't block a plan because it isn't exactly how you would have structured it. If it's executable and well-organized, approve it.
review-plan flow loads this skill to perform
the review of a plan.review-plan flow — never load this
skill directly for that. This is the criteria reference, not the flow.Do NOT use for: Single-file changes with obvious scope, or when the task is trivial enough to just do.
Every plan gets evaluated across these dimensions:
Does the plan cover everything needed to implement successfully?
Missing any of these is a gap, not a nit.
Are the tasks well-defined and independently executable?
Is the plan structured so implementation flows correctly?
Are the hard parts acknowledged and mitigated?
Can an implementer actually execute this?
If the plan proposes code shapes, function signatures, data structures, or API designs, evaluate those proposals against code-review-criteria:
When to apply: Only when the plan includes specific code snippets, type definitions, API contracts, or function signatures. Plans that only describe "what" without showing "how" skip this axis.
When you flag a structural problem in a plan, propose the fix — not just the problem:
Prefer the remedy that makes the plan immediately actionable over one that just flags the gap.
Plans should be scoped to a single deliverable:
1–5 tasks → Good. A focused feature or bug fix.
6–10 tasks → Acceptable for a moderate feature.
11–15 tasks → Large. Consider splitting into phases.
15+ tasks → Too large. Split into multiple plans.What counts as "one plan": A self-contained set of changes that delivers a single coherent capability. If you can describe the goal in one sentence, it's one plan.
Label every comment with its severity so the author knows what's required vs optional:
| Prefix | Meaning | Author Action |
|---|---|---|
| (no prefix) | Required change | Must address before implementation starts |
| Critical: | Blocks implementation | Missing security consideration, data integrity risk, fundamentally wrong approach |
| Nit: | Minor, optional | Author may ignore — wording, formatting |
| Optional: / Consider: | Suggestion | Worth considering but not required |
| FYI | Informational only | No action needed — context for future reference |
Lead with what matters. Order findings by leverage: missing risks and wrong sequencing first, then task quality gaps, then completeness, then nits. If you have one critical sequencing problem and ten nits, the sequencing problem is the review.
Before evaluating structure, understand intent:
- What is this plan trying to accomplish?
- What problem does it solve?
- What does "done" look like?Scan for missing sections before diving into content:
- Context present?
- Affected modules listed?
- Architecture decisions documented?
- Risks acknowledged?
- Testing strategy defined?
- Verification commands explicit?Walk through each task:
For each task:
1. Can I tell exactly what to build?
2. Are acceptance criteria specific and testable?
3. Is the size reasonable (not XL)?
4. Are dependencies clear?
5. Would I know which files to touch?Check the dependency graph:
- Are foundations built first?
- Does each task leave the system working?
- Are checkpoints placed correctly?
- Are high-risk items early?
- Is it vertically sliced?Put yourself in the implementer's shoes:
- Could I pick up task 1 and start coding without asking any questions?
- Are the verification commands copy-pasteable?
- Are file paths and function names specific?
- Is existing code referenced where I'd need to read it?Check that the plan can actually confirm it worked:
- What tests should pass after implementation?
- What build/compile commands are relevant?
- What manual checks are needed?
- How do we know the feature works end-to-end?If the plan includes code snippets, types, or API designs:
- Load code-review-criteria skill for criteria
- Check proposed signatures for edge cases
- Verify naming follows project conventions
- Confirm abstractions follow existing patterns
- Scan for security vectors in proposed APIs
- Check for performance issues in proposed data structures## Review: [Plan title]
### Completeness
- [ ] Context explains the problem and goal
- [ ] Affected modules are listed with paths
- [ ] Architecture decisions have rationale
- [ ] Testing strategy is defined
- [ ] Verification commands are explicit and project-specific
- [ ] Open questions are listed
### Task Quality
- [ ] Every task has acceptance criteria
- [ ] Every task has verification steps
- [ ] Tasks are sized XS–M (L acceptable, XL must be split)
- [ ] Task dependencies are stated
- [ ] Files likely touched are listed
### Architecture & Sequencing
- [ ] Order follows dependency graph (foundations first)
- [ ] Vertically sliced (not horizontal layers)
- [ ] Each task leaves system working
- [ ] Checkpoints exist between phases
- [ ] High-risk tasks are early
### Risk Coverage
- [ ] Edge cases identified
- [ ] Breaking changes / migrations noted
- [ ] Security implications considered
- [ ] Performance implications considered
- [ ] Rollback strategy exists (if applicable)
### Actionability
- [ ] File paths are specific
- [ ] Verification commands are copy-pasteable
- [ ] Existing code to read is referenced
- [ ] Conventions and patterns are noted
### Proposed Code Quality *(if plan includes implementation details)*
- [ ] Proposed types/signatures handle edge cases
- [ ] Proposed names follow project conventions
- [ ] Proposed abstractions follow existing patterns
- [ ] No security vectors in proposed APIs
- [ ] No performance issues in proposed structures
### Verdict
- [ ] **Approve** — Ready to implement
- [ ] **Request changes** — Gaps must be addressed| Rationalization | Reality |
|---|---|
| "I'll figure out the details during implementation" | That's how you discover blocking dependencies mid-task. Surface them now. |
| "The tasks are obvious, no need for criteria" | Write them anyway. Explicit criteria surface hidden assumptions. |
| "It's just a small feature, it doesn't need a plan" | Small features have edge cases too. 3 tasks with criteria takes 5 minutes. |
| "The plan is good enough" | "Good enough" without acceptance criteria means the implementer defines "done" — and they might define it differently. |
| "I'll add verification steps later" | Later never comes. The plan is the contract — define verification now. |
| "Risks are minimal" | Every change has risks. If you can't name them, you haven't thought about them. |
| "The file paths are obvious" | They're obvious to the author. The implementer might not know the codebase. |
| "The code in the plan is fine, it'll get reviewed later" | Plan-level code review catches design problems before implementation — fixing them after coding is more expensive. |
any/unknown/optional without justificationplanner skillcode-review-criteria — also the criteria source for axis 6security-and-hardeningtesting© penpot, MPL-2.0. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file
Just SKILL.md in .agents/skills/plan-review-criteria of penpot/penpot.
Open the folder on GitHubat commit 10955f1
Plan Review Criteria next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.
| Skill | Stars | Used in | Tokens | Auto-check | Licence | Repo updated |
|---|---|---|---|---|---|---|
| Plan Review Criteria this skillpenpot/penpot | 61k | — | ~3.3k | Automated safety check: Pass | MPL-2.0 | |
| ArenaJakeschincariol/arena-skill | 402 | — | ~4.8k | Automated safety check: Pass | MIT | |
| Dynamic WorkflowsHabitat-Thinking/ai-literacy-superpowers | 114 | — | ~1.9k | Automated safety check: Pass | Custom licence | |
| Woo AI Smokewoocommerce/woocommerce-ios | 358 | — | ~7.4k | Automated safety check: Notes | GPL-2.0 | |
| Subagent Driven DevelopmentAsvarox/allkaraoke | 261 | 38 repos | ~1.2k | Automated safety check: Pass | None | |
| Executing PlansGanyuanRan/Aegis | 1.3k | 1 repos | ~2.3k | Automated safety check: Pass | MIT |
Jakeschincariol/arena-skill
Make 100 versions of Claude fight to the death over one task.
Habitat-Thinking/ai-literacy-superpowers
This skill should be used when an agent is deciding whether to author a dynamic workflow — a self-authored, ephemeral multi-agent harness — for a task.
woocommerce/woocommerce-ios
Evaluate WooAIAssistant against a structured scenario suite with hard invariants + LLM-as-judge rubric scoring.
Asvarox/allkaraoke
A skill your agent uses when executing implementation plans with independent tasks in the current session
GanyuanRan/Aegis
A skill your agent uses when executing a written implementation plan across sessions or with review checkpoints.
umputun/cc-thingz
Execute plan tasks sequentially using subagents. An agent skill from umputun/cc-thingz.
penpot/penpot
Hardens code against vulnerabilities. An agent skill from penpot/penpot.
penpot/penpot
PR flow — open a new PR for the current task branch (validates base branch, commits, issue and push state) or update an existing PR's title or description to match Penpot conventions.
penpot/penpot
A cat clone with syntax highlighting, line numbers, and Git integration - a modern replacement for cat.
penpot/penpot
Run local CI-style checks with ./scripts/ci (lint, tests, format) per monorepo module.
penpot/penpot
Write or rewrite text in ASD-STE100 Simplified Technical English.
penpot/penpot
Code review criteria — the five review axes, core principles, severity format, and verdict for reviewing code changes.
Categories
Plan review criteria — the six review axes, severity rubric, approval standard, and output format for reviewing implementation plans. Plan Review Criteria is an agent skill from penpot/penpot. Plan review criteria — the six review axes, severity rubric, approval standard, and output format for reviewing implementation plans.
Plan Review Criteria fits situations like: tasks that involve Accessibility; tasks that involve Planning; tasks that involve Quizzes and assessments.
Run `npx skills add penpot/penpot --skill plan-review-criteria -a claude-code`. Or copy the skill folder (.agents/skills/plan-review-criteria in penpot/penpot) into .claude/skills/plan-review-criteria in your project. Claude Code loads it when a task matches its description.
Run `npx skills add penpot/penpot --skill plan-review-criteria -a codex`. Or copy the skill folder (.agents/skills/plan-review-criteria in penpot/penpot) into .agents/skills/plan-review-criteria in your project. Codex loads it when a task matches its description.
Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add penpot/penpot --skill plan-review-criteria -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/plan-review-criteria, .gemini/skills/plan-review-criteria, .github/skills/plan-review-criteria and .opencode/skills/plan-review-criteria in your project.
SKILL.md names no scripts, command-line tools or credentials: Plan Review Criteria is instructions for the agent only.
SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.
Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.
Plan Review Criteria is published under the MPL-2.0 licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.
About 3.3k tokens (SKILL.md is roughly 13k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.
Skills that share tags, products or a category with Plan Review Criteria: Arena (Jakeschincariol/arena-skill, 402 stars), Dynamic Workflows (Habitat-Thinking/ai-literacy-superpowers, 114 stars), Woo AI Smoke (woocommerce/woocommerce-ios, 358 stars) and Subagent Driven Development (Asvarox/allkaraoke, 261 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.
penpot (a GitHub organization) maintains it in penpot/penpot, which has 60,869 GitHub stars. The repository holds 24 skills in this directory. The repository was last updated on October 9, 2026.
Source: penpot/penpot on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.