Agent skill

Evaluate Artifact

by andreaskelm in andreaskelm/pm-brain

Evaluate quality of a PM artifact or substantive session output using the shared evaluation procedure — gut check, red/green flags, optional full scored review.

Custom licenceAuto-check passedProduct & Project Management

Install Evaluate Artifact

skills CLI
$ npx skills add andreaskelm/pm-brain --skill evaluate-artifact -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install andreaskelm/pm-brain evaluate-artifact --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/andreaskelm/pm-brain.git skills-src && mkdir -p .claude/skills && cp -r skills-src/.claude/skills/evaluate-artifact .claude/skills/evaluate-artifact && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
evaluate-artifact
GitHub stars
234
Token cost
~1.4k tokens
SKILL.md length
607 words
Files
1
Skills in repo
18
Repo updated
First seen
Licence
Custom licence

At a glance

Evaluate quality of a PM artifact or substantive session output using the shared evaluation procedure — gut check, red/green flags, optional full scored review.

  • Works in 8 steps: Identify what to evaluate → Map artifact → criteria file → While creating (lightweight, if still… → …
  • The user says evaluate this
  • SKILL.md covers Step 1 — Identify what to…, Step 2 — Map artifact →…, Step 3 — While creating… and Step 4 — Before calling it…, plus 6 more sections
  • Instructions only: no scripts, shell commands, URLs or credentials in SKILL.md

What it does

Evaluate Artifact is an agent skill from andreaskelm/pm-brain. Evaluate quality of a PM artifact or substantive session output using the shared evaluation procedure — gut check, red/green flags, optional full scored review. Use when the user says "evaluate this", "is this good enough", "quality check", "review the PRD/OKR/roadmap we just wrote", or after finishing an artifact. Includes end-of-session personal capture scan.

Its SKILL.md is about 1.4k tokens, which your agent loads only when the skill is triggered. It is a single SKILL.md file with no bundled scripts.

It sits in Product & Project Management, covering OKRs and executive reporting and PRD writing. The repository describes itself as: AI-powered product management thinking & operating system. Playbooks, guides, templates, and frameworks that bridge PM theory to daily execution.

When your agent uses it

  • The user says evaluate this
  • Is this good enough
  • Review the PRD/OKR/roadmap we just wrote
  • After finishing an artifact

Example prompts

  • “evaluate this”
  • “is this good enough”
  • “quality check”
  • “/evaluate-artifact”

Workflow steps

8 steps, taken from the step headings in SKILL.md.

  1. Identify what to evaluate
  2. Map artifact → criteria file
  3. While creating (lightweight, if still drafting)
  4. Before calling it done (quick check)
  5. Full evaluation (on request)
  6. Agent behavior (Level 2, if requested)
  7. Personal capture scan (always at session end)
  8. Log key findings (when doing Level 2 or meta)

What it can do on your machine

Read from SKILL.md and the folder at commit 38696ac. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    No scripts in the folder and no shell commands in SKILL.md.

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Evaluate Artifact loads about 1.4k tokens when it runs. Until then it costs about 95 tokens; SKILL.md has 607 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~95
When it runs · the whole SKILL.md, loaded when a task matches
~1.4k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

Its licence (Custom licence) doesn't allow us to republish the file, so here is its outline and opening line. It has 607 words (~1,373 tokens).

“One procedure for every artifact type. What changes is only the criteria file — red/green flags, weighted dimensions, antipatterns, rewrites. Process, math, and output format: system/EVALUATION.md (follow it; don't improvise scoring).”

— opening of SKILL.md by andreaskelm, Custom licence
name
evaluate-artifact

Read the full SKILL.md on GitHub

Files

Just SKILL.md in .claude/skills/evaluate-artifact of andreaskelm/pm-brain.

Open the folder on GitHubat commit 38696ac

Compare with similar skills

Evaluate Artifact next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Evaluate Artifact compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Evaluate Artifact this skillandreaskelm/pm-brain234—~1.4kAutomated safety check: PassCustom licence
Prd V09 Launch Metricsmattgierhart/PRD-driven-context-engineering180—~4.7kAutomated safety check: PassMIT
Utility Pm Criticproduct-on-purpose/pm-skills716—~1.5kAutomated safety check: PassApache-2.0
Notion Pmborghei/Claude-Skills891—~1.7kAutomated safety check: PassMIT
Prd V03 Features Value Planningmattgierhart/PRD-driven-context-engineering180—~2.1kAutomated safety check: PassMIT
Prd V03 Outcome Definitionmattgierhart/PRD-driven-context-engineering180—~1.9kAutomated safety check: PassMIT

Similar skills

  • Prd V09 Launch Metrics

    mattgierhart/PRD-driven-context-engineering

    Define success criteria and tracking setup for launch during PRD v0.9 Go-to-Market.

    180 GitHub stars~4.7k tokensUpdated 1 mo ago
    Product & Project ManagementAuto-check passed
  • Utility Pm Critic

    product-on-purpose/pm-skills

    Run adversarial review on a PM artifact via the pm-critic sub-agent.

    716 GitHub stars~1.5k tokensUpdated 3 days ago
    Product & Project ManagementAuto-check passed
  • Notion Pm

    borghei/Claude-Skills

    Notion expert for product management workflows. An agent skill from borghei/Claude-Skills.

    891 GitHub stars~1.7k tokensUpdated 4 days ago
    Product & Project ManagementAuto-check passed
  • Prd V03 Features Value Planning

    mattgierhart/PRD-driven-context-engineering

    Define and prioritize features with strategic traceability during PRD v0.3 Commercial Model.

    180 GitHub stars~2.1k tokensUpdated 1 mo ago
    Product & Project ManagementAuto-check passed
  • Prd V03 Outcome Definition

    mattgierhart/PRD-driven-context-engineering

    Define measurable success metrics (KPIs) tied to product type during PRD v0.3 Commercial Model.

    180 GitHub stars~1.9k tokensUpdated 1 mo ago
    Product & Project ManagementAuto-check passed
  • Prd V04 User Journey Mapping

    mattgierhart/PRD-driven-context-engineering

    Map user missions from trigger to value moment, organizing features into coherent paths during PRD v0.4 User Journeys.

    180 GitHub stars~2.4k tokensUpdated 1 mo ago
    Product & Project ManagementAuto-check passed

More from andreaskelm/pm-brain

All 18 skills in this repo
  • AI Product Management

    andreaskelm/pm-brain

    Ship and spec AI features, LLM products, agents, copilots, and generative UX — including when to use a model vs.

    234 GitHub stars~1.8k tokensUpdated yesterday
    Auto-check passed
  • Discovery Synthesis

    andreaskelm/pm-brain

    Plan customer discovery, turn interview snapshots into synthesis and evidence-based opportunities, build or update an Opportunity Solution Tree, map jobs and segments, and design RAT tests for the…

    234 GitHub stars~2.3k tokensUpdated yesterday
    Auto-check passed
  • Eng Design Collab

    andreaskelm/pm-brain

    Partner effectively with engineering and design — feasibility and scope negotiation, tech debt tradeoffs, design reviews, discovery with builders, and PRD handoffs that don't get thrown away.

    234 GitHub stars~2.1k tokensUpdated yesterday
    Auto-check passed
  • Experimentation

    andreaskelm/pm-brain

    Design and run product experiments at a practical PM level — A/B tests, hypothesis tests, rollouts, feature flags, and reading results without pretending to be a statistician.

    234 GitHub stars~2.3k tokensUpdated yesterday
    Auto-check passed
  • Launch Gtm

    andreaskelm/pm-brain

    Plan, tighten, or review a product launch and go-to-market motion: rollout phases, beta and GA readiness, sales enablement, marketing launch, and release comms (internal and external).

    234 GitHub stars~2.1k tokensUpdated yesterday
    Auto-check passed
  • North Star

    andreaskelm/pm-brain

    Define, sharpen, or audit a North Star metric and its input metrics tree, and decide which product metrics actually matter (leading vs.

    234 GitHub stars~2.5k tokensUpdated yesterday
    Auto-check passed

Questions about Evaluate Artifact

What does Evaluate Artifact do?

Evaluate quality of a PM artifact or substantive session output using the shared evaluation procedure — gut check, red/green flags, optional full scored review. Evaluate Artifact is an agent skill from andreaskelm/pm-brain. Evaluate quality of a PM artifact or substantive session output using the shared evaluation procedure — gut check, red/green flags, optional full scored review.

When should I use Evaluate Artifact?

Evaluate Artifact fits situations like: the user says evaluate this; is this good enough; review the PRD/OKR/roadmap we just wrote; after finishing an artifact.

How do I install Evaluate Artifact in Claude Code?

Run `npx skills add andreaskelm/pm-brain --skill evaluate-artifact -a claude-code`. Or copy the skill folder (.claude/skills/evaluate-artifact in andreaskelm/pm-brain) into .claude/skills/evaluate-artifact in your project. Claude Code loads it when a task matches its description.

How do I install Evaluate Artifact in Codex?

Run `npx skills add andreaskelm/pm-brain --skill evaluate-artifact -a codex`. Or copy the skill folder (.claude/skills/evaluate-artifact in andreaskelm/pm-brain) into .agents/skills/evaluate-artifact in your project. Codex loads it when a task matches its description.

Can I use Evaluate Artifact in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add andreaskelm/pm-brain --skill evaluate-artifact -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/evaluate-artifact, .gemini/skills/evaluate-artifact, .github/skills/evaluate-artifact and .opencode/skills/evaluate-artifact in your project.

What does Evaluate Artifact need to run?

SKILL.md names no scripts, command-line tools or credentials: Evaluate Artifact is instructions for the agent only.

Does Evaluate Artifact access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is Evaluate Artifact safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Evaluate Artifact use?

Evaluate Artifact has a licence file (the repository's licence) that doesn't match a standard licence. Read it on GitHub before reusing the skill.

How many tokens does Evaluate Artifact use?

About 1.4k tokens (SKILL.md is roughly 5.5k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Evaluate Artifact?

Skills that share tags, products or a category with Evaluate Artifact: Prd V09 Launch Metrics (mattgierhart/PRD-driven-context-engineering, 180 stars), Utility Pm Critic (product-on-purpose/pm-skills, 716 stars), Notion Pm (borghei/Claude-Skills, 891 stars) and Prd V03 Features Value Planning (mattgierhart/PRD-driven-context-engineering, 180 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Evaluate Artifact?

andreaskelm (a GitHub user) maintains it in andreaskelm/pm-brain, which has 234 GitHub stars. The repository holds 18 skills in this directory. The repository was last updated on October 9, 2026.

Source: andreaskelm/pm-brain on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.