Agent skill

Tool Procurement Eval

by mohitagw15856 in mohitagw15856/pm-claude-skills

Evaluate a new tool before it joins the stack — the problem-first framing (tools answer needs, not demos), the trial designed with success criteria upfront, the stack-fit check (integration…

MITAuto-check passedBusiness, Finance & HR

Install Tool Procurement Eval

skills CLI
$ npx skills add mohitagw15856/pm-claude-skills --skill tool-procurement-eval -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install mohitagw15856/pm-claude-skills tool-procurement-eval --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/mohitagw15856/pm-claude-skills.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/tool-procurement-eval .claude/skills/tool-procurement-eval && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
tool-procurement-eval
GitHub stars
1.4k
Token cost
~1.6k tokens
SKILL.md length
805 words
Files
1
Skills in repo
1,348
Repo updated
First seen
Licence
MIT

At a glance

Evaluate a new tool before it joins the stack — the problem-first framing (tools answer needs, not demos), the trial designed with success criteria upfront, the stack-fit check (integration…

  • Works in 5 steps: Need before tool: the statement —… → The overlap audit runs early and… → Trials have pre-written criteria and a… → …
  • Asked should we buy this tool
  • SKILL.md covers What This Skill Produces, Required Inputs, Framework: The Eval Rules and Output Format, plus 7 more sections
  • Instructions only: no scripts, shell commands, URLs or credentials in SKILL.md

What it does

Tool Procurement Eval is an agent skill from mohitagw15856/pm-claude-skills. Evaluate a new tool before it joins the stack — the problem-first framing (tools answer needs, not demos), the trial designed with success criteria upfront, the stack-fit check (integration, overlap, the tool-sprawl tax), and the security/data review sized to the stakes. Use when asked should we buy this tool, evaluate this software for the team, we have three tools that do this already, or run a proper trial before committing. Produces the need statement, the trial design with pre-set criteria, the stack-fit…

Its SKILL.md is about 1.6k tokens, which your agent loads only when the skill is triggered. It is a single SKILL.md file with no bundled scripts.

It sits in Business, Finance & HR, covering Vendor and procurement management. The repository describes itself as: 1255 professional Agent Skills for Claude, ChatGPT, Gemini, Cursor & Codex — PRDs, postmortems, leases, medical bills, layoffs, go-bags, new countries. Plain markdown, MIT, in… The licence is MIT.

When your agent uses it

  • Asked should we buy this tool
  • Evaluate this software for the team
  • We have three tools that do this already
  • Run a proper trial before committing

Example prompts

  • “/tool-procurement-eval”

Workflow steps

5 steps, taken from the first numbered list in SKILL.md.

  1. Need before tool: the statement — problem, owner, frequency, cost of the status quo — written before any vendor contact; requirements…
  2. The overlap audit runs early and honestly: does an owned tool cover 80% of the need? (The honest answer kills the request — cheaply and…
  3. Trials have pre-written criteria and a skeptic: "success = the team's weekly report time drops below 2 hours, and 4 of 6 pilots choose to…
  4. The security gate scales to the data: public-content tools get the light pass (vendor's security page, the data-processing basics)…
  5. The verdict gets logged either way: adopt → owner named, rollout planned, the renewal row created at signature ([the intake rule]) ·…

What it can do on your machine

Read from SKILL.md and the folder at commit 1cbf1f0. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    No scripts in the folder and no shell commands in SKILL.md.

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Tool Procurement Eval loads about 1.6k tokens when it runs. Until then it costs about 148 tokens; SKILL.md has 805 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~148
When it runs · the whole SKILL.md, loaded when a task matches
~1.6k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from mohitagw15856/pm-claude-skills at commit 1cbf1f0, republished under its MIT licence (© mohitagw15856). 805 words, ~1,616 tokens.

Download SKILL.mdSave it as .claude/skills/tool-procurement-eval/SKILL.md (or your agent's skills folder).
name
tool-procurement-eval
description
Evaluate a new tool before it joins the stack — the problem-first framing (tools answer needs, not demos), the trial designed with success criteria upfront, the stack-fit check (integration, overlap, the tool-sprawl tax), and the security/data review sized to the stakes. Use when asked should we buy this tool, evaluate this software for the team, we have three tools that do this already, or run a proper trial before committing. Produces the need statement, the trial design with pre-set criteria, the stack-fit audit, and the adopt/decline verdict with its reasoning.

Tool Procurement Eval Skill

Tools enter stacks backwards: someone sees a demo, gets excited, and the "evaluation" becomes a justification ritual (vendor-comparison-matrix fights this at the compare stage; this skill fights it at the door). The forward order: the need stated first (which problem, whose, costing what — purchase-justification arithmetic), the stack-fit check before the trial (does something we own already do this? — the overlap audit that kills half of tool requests honestly), the trial designed with success criteria written before day one (or the trial's warm feelings decide), and the security/data review sized to what the tool touches — because the fun tool that ingests customer data is a compliance decision wearing a productivity costume.

What This Skill Produces

  • The need statement — the problem, its owner, its cost, and the requirements it implies (must vs. nice — pre-demo)
  • The stack-fit audit — the overlap check against owned tools, the integration reality, and the sprawl tax named
  • The trial design — duration, participants, the pre-written success criteria, and the decision date
  • The verdict — adopt (with owner and rollout) / decline (with the reason logged) — plus the security-review gate where data warrants

Required Inputs

Ask for these if not provided:

  • The problem, not the tool — what's broken/slow/manual today, for whom, costing what; "I saw this cool tool" gets reverse-engineered into its implied need, which sometimes evaporates on contact
  • The current stack — what's owned that's adjacent (the overlap audit needs the inventory — contract-renewal-tracker's list is the source); most orgs own 30% more capability than they use
  • What the tool would touch — customer data? Credentials? Just public content? The security review's depth follows (skill-vetting blast-radius thinking, applied to SaaS)
  • The trial population — who'd actually test it (the enthusiast and a skeptic — enthusiast-only trials always pass)

Framework: The Eval Rules

  1. Need before tool: the statement — problem, owner, frequency, cost of the status quo — written before any vendor contact; requirements derive from it (must-haves gate, nice-to-haves score). Tools without a need statement are solutions shopping for problems on your budget.
  2. The overlap audit runs early and honestly: does an owned tool cover 80% of the need? (The honest answer kills the request — cheaply and correctly.) Half-used owned tools get their config/training gap named instead ("we own this in Notion; nobody set it up") — the sprawl tax (another login, another admin, another renewal row, another data silo) is a real cost the shiny demo never quotes.
  3. Trials have pre-written criteria and a skeptic: "success = the team's weekly report time drops below 2 hours, and 4 of 6 pilots choose to keep it" — written before day one (survey-design-basics pre-commitment), tested with the enthusiast and the skeptic (the enthusiast finds the ceiling; the skeptic finds the floor), timeboxed with a decision date. Trials without criteria are extended demos that always end in purchase.
  4. The security gate scales to the data: public-content tools get the light pass (vendor's security page, the data-processing basics); anything touching customer data, credentials, or internal documents gets the real review (where's the data stored, who can access, the deletion story, the tos-decoder read of their terms) — before the trial pipes real data in, not after. "It's just a trial" is how customer data ends up in un-reviewed vendors.
  5. The verdict gets logged either way: adopt → owner named, rollout planned, the renewal row created at signature ([the intake rule]) · decline → the reason in the decision-log ("evaluated [tool] July 2026 — declined: 80% covered by owned stack") — because the same tool returns with a new champion every eighteen months, and the log converts the rematch into a lookup.
Show full SKILL.md (215 more words)Show less

Output Format

Tool Eval: [tool] — need: [the problem statement]

The Need + Requirements

[Problem/owner/cost · must-haves · nice-to-haves — dated pre-demo]

Stack-Fit Audit

[Overlap: (owned tools × coverage %) · the config-gap finding if applicable · integration reality · the sprawl tax lines]

Trial Design

[Duration · pilots (enthusiast + skeptic named) · the pre-written criteria · decision date · the security gate status before real data]

The Verdict

[Adopt: owner/rollout/renewal-row · Decline: the logged reason · either way: in the decision log]

Quality Checks

  • The need statement predates vendor contact
  • The overlap audit ran against the real inventory with honest coverage
  • Trial criteria were written before day one and include a skeptic
  • The security review preceded real data entering the trial
  • The verdict is logged with reasons, adopt or decline

Anti-Patterns

  • Do not evaluate backwards from the demo — the need statement is the eval's spine
  • Do not skip the overlap audit — the cheapest tool is the one already owned and unconfigured
  • Do not run criteria-free trials — warm feelings always vote adopt
  • Do not pipe customer data into "just a trial" — the gate runs first at exactly that moment
  • Do not decline silently — the unlogged rejection is next year's rematch, at full cost

Example Trigger Phrases

  • "Should we buy this tool?"
  • "Evaluate this software for the team."
  • "We have three tools that do this already."
  • "Run a proper trial before committing."

© mohitagw15856, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

Just SKILL.md in skills/tool-procurement-eval of mohitagw15856/pm-claude-skills.

Open the folder on GitHubat commit 1cbf1f0

Compare with similar skills

Tool Procurement Eval next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Tool Procurement Eval compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Tool Procurement Eval this skillmohitagw15856/pm-claude-skills1.4k—~1.6kAutomated safety check: PassMIT
Serenity Alphahaskaomni/serenity-skill633—~2.6kAutomated safety check: PassMIT
Scorecard Matrixpnp/sharepoint-skills133—~1.9kAutomated safety check: PassMIT
Oma Imagefirst-fluke/oh-my-agent1.3k—~2kAutomated safety check: PassMIT
Buyer Job Intent Analysiselvisun/newsjack1.5k—~1.4kAutomated safety check: PassMIT
Master Builderibuilder/massing122—~2.6kAutomated safety check: PassMIT

Similar skills

  • Serenity Alpha

    haskaomni/serenity-skill

    Translate market-moving news into investable alpha hypotheses by mapping observed demand changes to revenue lines, supply chains, small-cap financial elasticity, market misclassification, validation…

    633 GitHub stars~2.6k tokensUpdated 2 mo ago
    Business, Finance & HRAuto-check passed
  • Scorecard Matrix

    pnp/sharepoint-skills

    Generates a polished, self-contained HTML heatmap scorecard — a weighted comparison matrix where entities (rows) are scored across dimensions (columns), with computed totals, rank badges, and a…

    133 GitHub stars~1.9k tokensUpdated yesterday
    Business, Finance & HRAuto-check passed
  • Oma Image

    first-fluke/oh-my-agent

    Generate raster images or reference-guided variations through the OMA image CLI.

    1.3k GitHub stars~2k tokensUpdated today
    Business, Finance & HRAuto-check passed
  • Recover source-bound buyer jobs, struggling moments, desired progress, forces, workarounds, information acts, journey states, criteria, constraints, roles, locales, and authentic language.

    1.5k GitHub stars~1.4k tokensUpdated 3 days ago
    Business, Finance & HRAuto-check passed
  • Master Builder

    ibuilder/massing

    Reason like a master builder — one mind holding an entire built-asset project from raw land through design, construction, handover, operations, and disposition, anywhere in the world.

    122 GitHub stars~2.6k tokensUpdated yesterday
    Business, Finance & HRAuto-check passed
  • Energy Procurement

    affaan-m/ECC

    Procure electricity and natural gas for commercial and industrial facilities: tariff and rate-schedule optimization, demand-charge mitigation, supplier RFPs, fixed/index/block-and-index hedging…

    276k GitHub starsUsed in 4 repos~7.4k tokens
    Business, Finance & HRAuto-check passed

More from mohitagw15856/pm-claude-skills

All 1,348 skills in this repo
  • Car Tco

    mohitagw15856/pm-claude-skills

    Compare the total cost of car ownership across buy-new, buy-used, lease, and keep-your-current-car — depreciation, insurance, maintenance ramp, and fuel over a real horizon, not just the monthly…

    1.4k GitHub stars~1.1k tokensUpdated yesterday
    Auto-check passed
  • Cs Health Scorecard

    mohitagw15856/pm-claude-skills

    Build a customer health scorecard for a specific account. An agent skill from mohitagw15856/pm-claude-skills.

    1.4k GitHub stars~2.4k tokensUpdated yesterday
    Auto-check passed
  • Exit Waterfall

    mohitagw15856/pm-claude-skills

    Compute who gets what at each exit price from a cap table — liquidation preferences, conversion points, and where the founders' share collapses.

    1.4k GitHub stars~1.1k tokensUpdated yesterday
    Auto-check passed
  • Feature Prioritisation

    mohitagw15856/pm-claude-skills

    Apply prioritisation frameworks (RICE, MoSCoW, Kano, ICE, Opportunity Scoring) to rank features and backlog items.

    1.4k GitHub stars~2k tokensUpdated yesterday
    Auto-check passed
  • Fire Number

    mohitagw15856/pm-claude-skills

    Compute a financial-independence (FIRE) target and years-to-reach with every assumption labeled as an assumption — plus a sensitivity table instead of a single false-precision answer.

    1.4k GitHub stars~1.1k tokensUpdated yesterday
    Auto-check passed
  • Freelance Rate

    mohitagw15856/pm-claude-skills

    Derive a freelance day/hourly rate backwards from target income, honest billable utilization, overhead, and the self-employment tax premium — the arithmetic that proves a rate is not salary÷2000.

    1.4k GitHub stars~1.2k tokensUpdated yesterday
    Auto-check passed

Questions about Tool Procurement Eval

What does Tool Procurement Eval do?

Evaluate a new tool before it joins the stack — the problem-first framing (tools answer needs, not demos), the trial designed with success criteria upfront, the stack-fit check (integration…. Tool Procurement Eval is an agent skill from mohitagw15856/pm-claude-skills. Evaluate a new tool before it joins the stack — the problem-first framing (tools answer needs, not demos), the trial designed with success criteria upfront, the stack-fit check (integration, overlap, the tool-sprawl tax), and the security/data review sized to the stakes.

When should I use Tool Procurement Eval?

Tool Procurement Eval fits situations like: asked should we buy this tool; evaluate this software for the team; we have three tools that do this already; run a proper trial before committing.

How do I install Tool Procurement Eval in Claude Code?

Run `npx skills add mohitagw15856/pm-claude-skills --skill tool-procurement-eval -a claude-code`. Or copy the skill folder (skills/tool-procurement-eval in mohitagw15856/pm-claude-skills) into .claude/skills/tool-procurement-eval in your project. Claude Code loads it when a task matches its description.

How do I install Tool Procurement Eval in Codex?

Run `npx skills add mohitagw15856/pm-claude-skills --skill tool-procurement-eval -a codex`. Or copy the skill folder (skills/tool-procurement-eval in mohitagw15856/pm-claude-skills) into .agents/skills/tool-procurement-eval in your project. Codex loads it when a task matches its description.

Can I use Tool Procurement Eval in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add mohitagw15856/pm-claude-skills --skill tool-procurement-eval -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/tool-procurement-eval, .gemini/skills/tool-procurement-eval, .github/skills/tool-procurement-eval and .opencode/skills/tool-procurement-eval in your project.

What does Tool Procurement Eval need to run?

SKILL.md names no scripts, command-line tools or credentials: Tool Procurement Eval is instructions for the agent only.

Does Tool Procurement Eval access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is Tool Procurement Eval safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Tool Procurement Eval use?

Tool Procurement Eval is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Tool Procurement Eval use?

About 1.6k tokens (SKILL.md is roughly 6.5k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Tool Procurement Eval?

Skills that share tags, products or a category with Tool Procurement Eval: Serenity Alpha (haskaomni/serenity-skill, 633 stars), Scorecard Matrix (pnp/sharepoint-skills, 133 stars), Oma Image (first-fluke/oh-my-agent, 1.3k stars) and Buyer Job Intent Analysis (elvisun/newsjack, 1.5k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Tool Procurement Eval?

mohitagw15856 (a GitHub user) maintains it in mohitagw15856/pm-claude-skills, which has 1,434 GitHub stars. The repository holds 1,348 skills in this directory. The repository was last updated on October 9, 2026.

Source: mohitagw15856/pm-claude-skills on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.