Agent skill

Foundry Failure Triage

by aviggiano in aviggiano/security

Classify failing Foundry fuzz, property, invariant, and differential tests.

MITAuto-check passed

Install Foundry Failure Triage

skills CLI
$ npx skills add aviggiano/security --skill foundry-failure-triage -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install aviggiano/security foundry-failure-triage --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/aviggiano/security.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/foundry-failure-triage .claude/skills/foundry-failure-triage && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
foundry-failure-triage
GitHub stars
144
Token cost
~637 tokens
SKILL.md length
305 words
Files
2
Skills in repo
10
Repo updated
First seen
Licence
MIT

At a glance

Classify failing Foundry fuzz, property, invariant, and differential tests.

  • Works in 6 steps: Read AGENTS.md and any testing-plan docs. → Run the requested Foundry command and… → Reproduce each failing test with the… → …
  • Codex needs to run forge tests
  • SKILL.md covers Workflow, Categories, Strictness and Subagent Prompt Shape, plus 1 more section
  • Instructions only: no scripts, shell commands, URLs or credentials in SKILL.md

What it does

Foundry Failure Triage is an agent skill from aviggiano/security. Classify failing Foundry fuzz, property, invariant, and differential tests. Use when Codex needs to run forge tests, reproduce counterexamples, spawn fresh investigations, separate harness defects from reference bugs and production bugs, and report strict-equality failures without masking them.

Its SKILL.md is about 640 tokens, which your agent loads only when the skill is triggered. The skill folder holds 2 other files (for example `agents/openai.yaml`).

The repository describes itself as: Security Reviews and Audit Checklists. The licence is MIT.

When your agent uses it

  • Codex needs to run forge tests
  • Reproduce counterexamples
  • Spawn fresh investigations
  • Separate harness defects from reference bugs and production bugs

Example prompts

  • “/foundry-failure-triage”

Workflow steps

6 steps, taken from the first numbered list in SKILL.md.

  1. Read AGENTS.md and any testing-plan docs.
  2. Run the requested Foundry command and capture the failing test list.
  3. Reproduce each failing test with the narrowest command and useful verbosity.
  4. For independent failures, use fresh context or subagents when available.
  5. Classify each failure before fixing anything.
  6. Report a table before making changes unless the user has already authorized specific fixes.

What it can do on your machine

Read from SKILL.md and the folder at commit e18ce7d. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    No scripts in the folder and no shell commands in SKILL.md.

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Foundry Failure Triage loads about 637 tokens when it runs. Until then it costs about 80 tokens; SKILL.md has 305 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~80
When it runs · the whole SKILL.md, loaded when a task matches
~637

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from aviggiano/security at commit e18ce7d, republished under its MIT licence (© aviggiano). 305 words, ~637 tokens.

Download SKILL.mdSave it as .claude/skills/foundry-failure-triage/SKILL.md (or your agent's skills folder). This skill also uses 1 other file; get the full folder from GitHub.
name
foundry-failure-triage
description
Classify failing Foundry fuzz, property, invariant, and differential tests. Use when Codex needs to run forge tests, reproduce counterexamples, spawn fresh investigations, separate harness defects from reference bugs and production bugs, and report strict-equality failures without masking them.

Foundry Failure Triage

Workflow

  1. Read AGENTS.md and any testing-plan docs.
  2. Run the requested Foundry command and capture the failing test list.
  3. Reproduce each failing test with the narrowest command and useful verbosity.
  4. For independent failures, use fresh context or subagents when available.
  5. Classify each failure before fixing anything.
  6. Report a table before making changes unless the user has already authorized specific fixes.

Categories

  • harness defect: invalid setup, impossible bounds, asymmetric production/reference state, wrong ABI shape, stale interface, or test expectation not actually implied by public behavior.
  • reference bug: the independent model is incomplete or diverges from the public spec/interface.
  • production bug: production diverges from the public spec/interface or from a valid independent reference.
  • spec mismatch: the test encodes a desired/spec behavior that production does not satisfy and should remain failing for developer attention.
  • unknown: insufficient evidence; preserve the failing test.

Do not assume production is correct. Do not assume the reference is correct. Do not assume the test is correct.

Strictness

Treat one-unit and one-wei differences as real findings unless the spec defines tolerance. Never relax equality only because a failure is small.

Only fix tests when confidence is high that the test would otherwise create a false positive. If confidence is low or medium, keep the failing test and explain the ambiguity.

Subagent Prompt Shape

Give each fresh investigation:

  • failing test name and file
  • exact forge command or counterexample
  • relevant public interface/spec paths or URLs
  • classification categories
  • instruction not to edit files unless asked

Ask for root cause, classification, confidence, and recommendation. Merge results into one table and resolve conflicts by doing a fresh local check.

Report Table

Use columns like:

Failing testCategoryRoot causeRecommendation

When the user needs developer-facing output, include exact file/test names, enough evidence to reproduce, and whether the test should remain failing.

© aviggiano, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

SKILL.md and 1 other file in skills/foundry-failure-triage of aviggiano/security.

  • SKILL.md
  • agents/openai.yaml

Open the folder on GitHubat commit e18ce7d

Compare with similar skills

Foundry Failure Triage next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Foundry Failure Triage compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Foundry Failure Triage this skillaviggiano/security144—~637Automated safety check: PassMIT
GreptimeDB Fuzz CI Failure InvestigationGreptimeTeam/greptimedb6.7k—~4.4kAutomated safety check: PassApache-2.0
Nemoclaw Maintainer Classify CI FailureNVIDIA/NemoClaw23k—~806Automated safety check: PassApache-2.0
Fuzzmillionco/react-doctor15k—~839Automated safety check: PassCustom licence
cargo-fuzz Rust Fuzzingtrailofbits/skills7.5k—~2.9kAutomated safety check: PassCC-BY-SA-4.0
CSS At Propertythedaviddias/Front-End-Checklist74k—~602Automated safety check: PassMIT

Similar skills

  • Diagnoses a failed GreptimeDB fuzz CI job by pulling its GitHub Actions logs and fuzz artifacts, then matching the evidence to the local source code.

    6.7k GitHub stars~4.4k tokensUpdated today
    Testing & QAAuto-check passed
  • Classify one failed NemoClaw GitHub Actions job using bounded, redacted logs and optional retained artifacts.

    23k GitHub stars~806 tokensUpdated today
    Testing & QAAuto-check passed
  • Fuzz

    millionco/react-doctor

    Fuzz React Doctor rules for crashes, slowness, false positives, and mutation-sensitive diagnostics with @react-doctor/fuzz.

    15k GitHub stars~839 tokensUpdated today
    DevelopmentAuto-check passed
  • cargo-fuzz Rust Fuzzing

    trailofbits/skills

    Official

    Sets up cargo-fuzz for a Cargo-based Rust project: nightly toolchain, fuzz targets, structured inputs, sanitizers, coverage and reproducing crashes.

    7.5k GitHub stars~2.9k tokensUpdated yesterday
    SecurityAuto-check passed
  • CSS At Property

    thedaviddias/Front-End-Checklist

    A skill your agent uses when implementing animated gradients, complex CSS transitions that involve custom property values, or building a typed design token system where custom property misuse should…

    74k GitHub stars~602 tokensUpdated 4 days ago
    Frontend & DesignAuto-check passed
  • Pester Failure Analysis

    PowerShell/PowerShell

    Investigates failing Pester tests in PowerShell CI jobs by following a six-step workflow from pull request status to documented fix recommendations.

    56k GitHub stars~5.1k tokensUpdated yesterday
    Testing & QAAuto-check passed

More from aviggiano/security

All 10 skills in this repo
  • Foundry Deploy Fixtures

    aviggiano/security

    Create or refactor Foundry deployment fixtures for Solidity tests.

    144 GitHub stars~634 tokensUpdated 25 days ago
    Auto-check passed
  • Foundry Fuzz Mirrors

    aviggiano/security

    Create Foundry fuzz tests from deterministic unit tests. An agent skill from aviggiano/security.

    144 GitHub stars~584 tokensUpdated 25 days ago
    Auto-check passed
  • Foundry Spec Properties

    aviggiano/security

    Turn whitepapers, protocol specs, and public documentation into Foundry property tests.

    144 GitHub stars~489 tokensUpdated 25 days ago
    Auto-check passed
  • Foundry Test Campaign

    aviggiano/security

    Master skill for running an end-to-end multi-pass Foundry testing campaign for Solidity projects.

    144 GitHub stars~1.6k tokensUpdated 25 days ago
    Auto-check passed
  • Stateful Invariant Testing

    aviggiano/security

    Build metric-driven Chimera/create-chimera-app stateful invariant testing campaigns for Solidity projects.

    144 GitHub stars~2.8k tokensUpdated 25 days ago
    Auto-check passed
  • Foundry Differential Tests

    aviggiano/security

    Create Foundry differential tests comparing production Solidity contracts against an independent reference model.

    144 GitHub stars~613 tokensUpdated 25 days ago
    Auto-check passed

Questions about Foundry Failure Triage

What does Foundry Failure Triage do?

Classify failing Foundry fuzz, property, invariant, and differential tests. Foundry Failure Triage is an agent skill from aviggiano/security. Classify failing Foundry fuzz, property, invariant, and differential tests.

When should I use Foundry Failure Triage?

Foundry Failure Triage fits situations like: Codex needs to run forge tests; reproduce counterexamples; spawn fresh investigations; separate harness defects from reference bugs and production bugs.

How do I install Foundry Failure Triage in Claude Code?

Run `npx skills add aviggiano/security --skill foundry-failure-triage -a claude-code`. Or copy the skill folder (skills/foundry-failure-triage in aviggiano/security) into .claude/skills/foundry-failure-triage in your project. Claude Code loads it when a task matches its description.

How do I install Foundry Failure Triage in Codex?

Run `npx skills add aviggiano/security --skill foundry-failure-triage -a codex`. Or copy the skill folder (skills/foundry-failure-triage in aviggiano/security) into .agents/skills/foundry-failure-triage in your project. Codex loads it when a task matches its description.

Can I use Foundry Failure Triage in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add aviggiano/security --skill foundry-failure-triage -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/foundry-failure-triage, .gemini/skills/foundry-failure-triage, .github/skills/foundry-failure-triage and .opencode/skills/foundry-failure-triage in your project.

What does Foundry Failure Triage need to run?

SKILL.md names no scripts, command-line tools or credentials: Foundry Failure Triage is instructions for the agent only.

Does Foundry Failure Triage access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is Foundry Failure Triage safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Foundry Failure Triage use?

Foundry Failure Triage is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Foundry Failure Triage use?

About 637 tokens (SKILL.md is roughly 2.5k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Foundry Failure Triage?

Skills that share tags, products or a category with Foundry Failure Triage: GreptimeDB Fuzz CI Failure Investigation (GreptimeTeam/greptimedb, 6.7k stars), Nemoclaw Maintainer Classify CI Failure (NVIDIA/NemoClaw, 23k stars), Fuzz (millionco/react-doctor, 15k stars) and cargo-fuzz Rust Fuzzing (trailofbits/skills, 7.5k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Foundry Failure Triage?

aviggiano (a GitHub user) maintains it in aviggiano/security, which has 144 GitHub stars. The repository holds 10 skills in this directory. The repository was last updated on September 15, 2026.

Source: aviggiano/security on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.