Agent skill

Test Audit

by yielded-dev in yielded-dev/agent

Audit existing tests for low value, implementation coupling, duplication, and test-only production seams; apply the value bar when reviewing test changes.

MITAuto-check passed

Install Test Audit

skills CLI
$ npx skills add yielded-dev/agent --skill test-audit -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install yielded-dev/agent test-audit --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/yielded-dev/agent.git skills-src && mkdir -p .claude/skills && cp -r skills-src/.agents/skills/test-audit .claude/skills/test-audit && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
test-audit
GitHub stars
137
Token cost
~1.8k tokens
SKILL.md length
901 words
Files
4
Skills in repo
3
Repo updated
First seen
Licence
MIT

At a glance

Audit existing tests for low value, implementation coupling, duplication, and test-only production seams; apply the value bar when reviewing test changes.

  • Works in 4 steps: What observable behavior, invariant, or… → What credible regression makes it fail? → Why does existing proof not catch that… → …
  • SKILL.md covers Authoring gate, Junk patterns, Discovery and retention and Candidate evidence, plus 2 more sections
  • Calls git

What it does

Test Audit is an agent skill from yielded-dev/agent. Audit existing tests for low value, implementation coupling, duplication, and test-only production seams; apply the value bar when reviewing test changes.

Its SKILL.md is about 1.8k tokens, which your agent loads only when the skill is triggered. The skill folder holds 3 other files (for example `CAMPAIGN.md`).

The repository describes itself as: Yielded Agent: an agent harness toolkit for TypeScript, built on Effect and Effect AI. The licence is MIT.

Example prompts

  • “/test-audit”

Workflow steps

4 steps, taken from the first numbered list in SKILL.md.

  1. What observable behavior, invariant, or independent contract does it protect?
  2. What credible regression makes it fail?
  3. Why does existing proof not catch that failure? Each contract has one primary owner at
  4. Does it require a production export, flag, wrapper, or injection hook with no production

What it can do on your machine

Read from SKILL.md and the folder at commit ba5813e. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Shell commands in SKILL.md call:

    • git

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    Links to these hosts (documentation or services it may open):

    • github.com

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Test Audit loads about 1.8k tokens when it runs. Until then it costs about 41 tokens; SKILL.md has 901 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~41
When it runs · the whole SKILL.md, loaded when a task matches
~1.8k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from yielded-dev/agent at commit ba5813e, republished under its MIT licence (© yielded-dev). 901 words, ~1,750 tokens.

Download SKILL.mdSave it as .claude/skills/test-audit/SKILL.md (or your agent's skills folder). This skill also uses 3 other files; get the full folder from GitHub.
name
test-audit
description
Audit existing tests for low value, implementation coupling, duplication, and test-only production seams; apply the value bar when reviewing test changes.
license
MIT

Test Audit

Adapted from OpenClaw's test-audit skill. See NOTICE and LICENSE.

Optimize for confidence, not deletion count. Audit mode discovers a few high-confidence candidates and reports evidence before editing. Authoring mode checks proposed tests at write time. For an explicitly requested whole-subsystem pruning campaign, read CAMPAIGN.md.

Read root and scoped AGENTS.md files and the repository's testing skill and test selection. Those policies govern whether new automation is justified; this skill does not authorize replacement tests, production changes, commits, publication, or deployment. Keep reports and investigation evidence outside the product repository, or in task/PR artifacts.

Authoring gate

Before adding or expanding a test, establish why automation is necessary under the testing policy, then answer:

  1. What observable behavior, invariant, or independent contract does it protect?
  2. What credible regression makes it fail?
  3. Why does existing proof not catch that failure? Each contract has one primary owner at the strongest boundary. Another layer needs a distinct transport, lifecycle, or adapter risk that the owner cannot exercise. Prefer an existing table or fixture to duplication.
  4. Does it require a production export, flag, wrapper, or injection hook with no production caller? Exercise the real boundary instead.

A missing answer means the test is not ready. Check every junk pattern below; retain a match only when it independently guards a named contract. Behavior-preserving refactoring should not break a behavior test. A regression must demonstrably fail on the pre-fix owner for the intended reason and pass after the repair. Keep one primary regression at its owner boundary.

Junk patterns

  • Assertion-free coverage probes, self-comparisons, and identity copiers.
  • Copied fixtures, inventories, manifests, export lists, or schema definitions.
  • Exact source, import, or string greps tied to incidental implementation.
  • Private predicates or call shapes already exercised at a real boundary.
  • Duplicate invocations of a contract, including adapter-local replays of shared helpers.
  • Tests that preserve test-only exports, globals, wrappers, or otherwise dead production code.
  • Expected values generated by the helper or renderer under test.
  • Mocks that implement the asserted behavior, or one identical mock for different APIs.
  • Fixtures that supply the receipt, admission, or callback ordering the owner must produce; persistence assertions against a store the production path never writes.
  • Capability tests that restate flags rather than exercise promised delivery or acknowledgement.
  • Negative controls that pass because of an unrelated guard or unreachable rejection path.
  • Names that promise more than the input and assertions exercise.

Discovery and retention

Keep discovery read-only. For broad scope, use independent owner-boundary lanes across core, capabilities/durability, storage, platform adapters, examples, tooling, and cross-cutting patterns. Use an available orchestrate skill when applicable; keep workers read-only with disjoint scopes and one primary integrator.

Before judging a candidate, read the complete test and production owner, entry point, callers, callees, sibling implementations, overlapping tests, CI routing, and relevant history. Inspect dependency source or types when a test claims dependency-backed behavior. Judge what the assertions can detect, not their names, length, citations, or test count.

Keep independent public API, protocol, config, migration, storage, security, platform, default, prompt-byte, generated cross-language, package, release, and architecture contracts when their failure would escape remaining checks and repeatable workflow verification is insufficient. Preserve narrow compile-time inference checks and adapter proof with distinct risks. Observable call ordering and credible regressions can also justify retention.

Source inspection can be the cheapest independent guard when it protects a user-facing key, byte, path, or authority contract and survives identifier-only refactoring. Static or slow is not a deletion reason. A retained baseline failure is a possible product bug: reproduce it and report or repair its owner within scope rather than deleting the test.

Show full SKILL.md (307 more words)Show less

Candidate evidence

Record each field before proposing an edit; a missing field means deletion is not ready:

  • Exact test name and location.
  • Failure the test can actually detect.
  • Non-test callers of the production or support seam.
  • Stronger remaining owner-boundary proof, or why no independent contract needs proof.
  • Relevant history and the reason the test or seam exists; label unknown history.
  • Production or test-support deletion unlocked, if any.
  • Risk and the focused Vite+ validation command.

Separate verified candidates, retained false positives, and unresolved hypotheses.

Authorized edits and validation

When cleanup is in scope, choose one coherent owner-boundary batch. Remove obsolete test-only seams and orphaned support instead of preserving aliases. Consolidate duplicate package or dependency assertions at their canonical owner. Preserve public and persisted contracts; uncertainty is not evidence for deletion. Do not write replacement tests that restate the same implementation or increase cleanup scope to inflate deletion counts.

Do not edit source or tests while their suites run in the same checkout. Follow repository command authority: use vp run -F <workspace> test for workspace suites or vp -C <package-directory> test <file-filter> for focused proof, consulting command help and runner configuration before filtering. Root vp test filename filters can also match tests inside .worktrees; scope the runner to its package when worktrees exist. For removed source greps, exercise the executable or dry-run that owns the contract. Run applicable formatting and git diff --check, then the required vp run ready handoff gate. Preserve blocked proof explicitly. Report production/tooling separately from tests and test support using the final diff. For PRs, use open-pull-request.

Handoff

Report actionable findings and their remaining proof, production simplifications available, retained false positives and reasons, focused/full proof actually run, blocked proof, and named follow-ups. If edits occurred, include production versus test LOC and PR/merge state. For read-only audits, say explicitly that tests and production source were unchanged.

© yielded-dev, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

SKILL.md and 3 other files in .agents/skills/test-audit of yielded-dev/agent.

  • SKILL.md
  • CAMPAIGN.md
  • LICENSE
  • NOTICE

Open the folder on GitHubat commit ba5813e

Compare with similar skills

Test Audit next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Test Audit compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Test Audit this skillyielded-dev/agent137—~1.8kAutomated safety check: PassMIT
Tag Duplicate PRs Issuesopenclaw/openclaw392k—~4kAutomated safety check: PassMIT
Duplicate JSthedaviddias/Front-End-Checklist74k—~423Automated safety check: PassMIT
Value Propositionphuryn/pm-skills27k—~1.5kAutomated safety check: PassMIT
Value Prop Statementsphuryn/pm-skills27k—~758Automated safety check: PassMIT
Duplicate Descriptionthedaviddias/Front-End-Checklist74k—~695Automated safety check: PassMIT

Similar skills

  • Tag Duplicate PRs Issues

    openclaw/openclaw

    Use gitcrawl to search duplicate OpenClaw PRs/issues, group related work in prtags, and sync duplicate state to GitHub.

    392k GitHub stars~4k tokensUpdated today
    Research & ScienceAuto-check passed
  • Duplicate JS

    thedaviddias/Front-End-Checklist

    A skill your agent uses when auditing slow page loads, heavy assets, or rendering delays related to Remove duplicate JavaScript libraries.

    74k GitHub stars~423 tokensUpdated yesterday
    Frontend & DesignAuto-check passed
  • Value Proposition

    phuryn/pm-skills

    Design a detailed value proposition using a 6-part JTBD template — Who, Why, What before, How, What after, Alternatives.

    27k GitHub stars~1.5k tokensUpdated 23 days ago
    Marketing & SEOAuto-check passed
  • Value Prop Statements

    phuryn/pm-skills

    Generate value proposition statements for marketing, sales, and onboarding from existing value propositions.

    27k GitHub stars~758 tokensUpdated 23 days ago
    Marketing & SEOAuto-check passed
  • Duplicate Description

    thedaviddias/Front-End-Checklist

    A skill your agent uses when auditing a site's meta tag uniqueness, generating page-specific meta descriptions, or reviewing CMS templates that inject the same description globally.

    74k GitHub stars~695 tokensUpdated yesterday
    Marketing & SEOAuto-check passed
  • Duplicate Sweep

    superset-sh/superset

    Find and merge duplicate Linear issues — group reports of the same underlying bug, pick the survivor, and move the evidence across.

    15k GitHub stars~841 tokensUpdated today
    Auto-check passed

More from yielded-dev/agent

  • Simplify

    yielded-dev/agent

    Simplify code or workflows, including explicitly requested pre-production removal of legacy compatibility.

    137 GitHub stars~486 tokensUpdated today
    Auto-check passed
  • Testing

    yielded-dev/agent

    Plan E2E verification and reproducible evidence, audit test value, or test an isolated system before implementation.

    137 GitHub stars~487 tokensUpdated today
    Auto-check passed

Questions about Test Audit

What does Test Audit do?

Audit existing tests for low value, implementation coupling, duplication, and test-only production seams; apply the value bar when reviewing test changes. Test Audit is an agent skill from yielded-dev/agent. Audit existing tests for low value, implementation coupling, duplication, and test-only production seams; apply the value bar when reviewing test changes.

How do I install Test Audit in Claude Code?

Run `npx skills add yielded-dev/agent --skill test-audit -a claude-code`. Or copy the skill folder (.agents/skills/test-audit in yielded-dev/agent) into .claude/skills/test-audit in your project. Claude Code loads it when a task matches its description.

How do I install Test Audit in Codex?

Run `npx skills add yielded-dev/agent --skill test-audit -a codex`. Or copy the skill folder (.agents/skills/test-audit in yielded-dev/agent) into .agents/skills/test-audit in your project. Codex loads it when a task matches its description.

Can I use Test Audit in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add yielded-dev/agent --skill test-audit -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/test-audit, .gemini/skills/test-audit, .github/skills/test-audit and .opencode/skills/test-audit in your project.

What does Test Audit need to run?

Going by SKILL.md and its folder, Test Audit needs the command-line tools its instructions call (git).

Does Test Audit access the network?

SKILL.md names 1 domain. As links in the text: github.com. This is read from the text; nothing was executed.

Is Test Audit safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Test Audit use?

Test Audit is published under the MIT licence (declared in SKILL.md). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Test Audit use?

About 1.8k tokens (SKILL.md is roughly 7k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Test Audit?

Skills that share tags, products or a category with Test Audit: Tag Duplicate PRs Issues (openclaw/openclaw, 392k stars), Duplicate JS (thedaviddias/Front-End-Checklist, 74k stars), Value Proposition (phuryn/pm-skills, 27k stars) and Value Prop Statements (phuryn/pm-skills, 27k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Test Audit?

yielded-dev (a GitHub organization) maintains it in yielded-dev/agent, which has 137 GitHub stars. The repository holds 3 skills in this directory. The repository was last updated on October 8, 2026.

Source: yielded-dev/agent on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.