Agent skill

Validation Harness Design

by hashgraph-online in hashgraph-online/awesome-codex-plugins

A skill your agent uses when designing repository validation commands, doctor scripts, test matrices, JSON or JUnit outputs, CI gates, smoke checks, or harness command surfaces.

Apache-2.0Auto-check passedTesting & QA

Install Validation Harness Design

skills CLI
$ npx skills add hashgraph-online/awesome-codex-plugins --skill validation-harness-design -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install hashgraph-online/awesome-codex-plugins validation-harness-design --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/hashgraph-online/awesome-codex-plugins.git skills-src && mkdir -p .claude/skills && cp -r skills-src/plugins/yfge/agent-harness-skills/skills/validation-harness-design .claude/skills/validation-harness-design && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
validation-harness-design
GitHub stars
1.2k
Token cost
~866 tokens
SKILL.md length
390 words
Files
3 (incl. references)
Skills in repo
686
Repo updated
First seen
Licence
Apache-2.0

At a glance

A skill your agent uses when designing repository validation commands, doctor scripts, test matrices, JSON or JUnit outputs, CI gates, smoke checks, or harness command surfaces.

  • Works in 7 steps: Search package.json, Makefile, scripts,… → Split validation into repo/docs,… → If there is no shared validation… → …
  • Designing repository validation commands
  • SKILL.md covers Overview, When To Use, Inputs Needed and Execution Order, plus 5 more sections
  • Instructions only: no scripts, shell commands, URLs or credentials in SKILL.md

What it does

Validation Harness Design is an agent skill from hashgraph-online/awesome-codex-plugins. Use when designing repository validation commands, doctor scripts, test matrices, JSON or JUnit outputs, CI gates, smoke checks, or harness command surfaces.

Its SKILL.md is about 870 tokens, which your agent loads only when the skill is triggered. The skill folder holds 3 other files, including reference files (for example `references/build-when-missing.md` and `references/environment-bootstrap.md`).

It sits in Testing & QA, covering Unit testing. It works with JUnit. The repository describes itself as: A curated list of awesome OpenAI Codex / ChatGPT plugins, skills, and resources. The 1 Codex Marketplace. See live plugins at: https://hol.org/plugins/best-codex-plugins. The licence is Apache-2.0.

When your agent uses it

  • Designing repository validation commands
  • Harness command surfaces

Example prompts

  • “/validation-harness-design”

Workflow steps

7 steps, taken from the first numbered list in SKILL.md.

  1. Search package.json, Makefile, scripts, CI workflows, test directories, and docs.
  2. Split validation into repo/docs, contracts, unit/type/lint, runtime, and external-dependency layers.
  3. If there is no shared validation entrypoint, bootstrap the minimum command surface from references/build-when-missing.md.
  4. For each change type, choose the minimum command and escalation condition.
  5. Design a unified entrypoint: check_repo_harness or doctor handles environment/structure, while focused commands test behavior.
  6. Design report output: stdout for humans, JSON/JUnit for CI and artifacts.
  7. Require skip/fallback reasons; do not describe degraded validation as full validation.

What it can do on your machine

Read from SKILL.md and the folder at commit 78497e5. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    No scripts in the folder and no shell commands in SKILL.md (its code samples are markdown).

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Validation Harness Design loads about 866 tokens when it runs, and up to ~1.6k if it reads all its reference files. Until then it costs about 46 tokens; SKILL.md has 390 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~46
When it runs · the whole SKILL.md, loaded when a task matches
~866
With references · SKILL.md plus every file in references/, read only if the agent opens them
~1.6k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from hashgraph-online/awesome-codex-plugins at commit 78497e5, republished under its Apache-2.0 licence (© hashgraph-online). 390 words, ~866 tokens.

Download SKILL.mdSave it as .claude/skills/validation-harness-design/SKILL.md (or your agent's skills folder). This skill also uses 2 other files; get the full folder from GitHub.
name
validation-harness-design
description
Use when designing repository validation commands, doctor scripts, test matrices, JSON or JUnit outputs, CI gates, smoke checks, or harness command surfaces.

Validation Harness Design

Overview

Define the smallest repeatable command surface that proves repository changes are safe enough for review.

This skill turns scattered checks into commands agents can run, CI can reuse, and failures can diagnose. For shared harness terms, see ../../references/harness-patterns.md; when validation commands are absent, use references/build-when-missing.md. For setup probes and CI-safe fallbacks, see references/environment-bootstrap.md.

When To Use

  • The user wants check_repo_harness.py, doctor.py, a test matrix, or CI gates.
  • Existing validation commands are scattered and agents do not know what to run for docs-only, interface, service, logic, or runtime changes.
  • JSON, JUnit, or artifact output is needed.

Inputs Needed

  • Project shape, main surfaces, existing test commands, and CI configuration.
  • Change types and the minimum gate for each type.
  • Whether runtime, packaged-environment, external-dependency, or manual checks are needed.

Execution Order

  • First: Inventory existing commands and CI to confirm what already works.
  • Then: Design a layered validation matrix and unified entrypoint.
  • Finally: Output the command table, report formats, CI wiring, and fallback rules.

Step-by-Step Process

  1. Search package.json, Makefile, scripts, CI workflows, test directories, and docs.
  2. Split validation into repo/docs, contracts, unit/type/lint, runtime, and external-dependency layers.
  3. If there is no shared validation entrypoint, bootstrap the minimum command surface from references/build-when-missing.md.
  4. For each change type, choose the minimum command and escalation condition.
  5. Design a unified entrypoint: check_repo_harness or doctor handles environment/structure, while focused commands test behavior.
  6. Design report output: stdout for humans, JSON/JUnit for CI and artifacts.
  7. Require skip/fallback reasons; do not describe degraded validation as full validation.
Show full SKILL.md (137 more words)Show less

Checks

  • Runnable: commands work in a clean checkout or documented setup.
  • Layered: docs-only work does not require heavy runtime checks, and runtime changes do not stop at static checks.
  • Output: failures identify case, path, command, and artifact.
  • CI: local commands and CI logic are consistent.
  • Cost: minimum gates are fast, and heavy gates trigger only on higher-risk changes.

Output Format

markdown
# Validation Harness Design

## Detected Mapping
- validation:
- runtime-evidence:
- CI:

## Change-Type Matrix
| Change type | Minimum check | Escalation |
| --- | --- | --- |

## Command Surface
-

## Report Outputs
-

## CI Gates
-

## Fallback / Skip Policy
-

Common Mistakes

  • Requiring the heaviest end-to-end check for every change, which causes agents to skip validation.
  • Providing only a command list without a change-type matrix.
  • Emitting unstable JSON/JUnit output that cannot be aggregated later.
  • Omitting fallback records, then claiming a degraded path was full runtime validation.

Example Prompts

  • "Design a check_repo_harness.py command surface for this repo."
  • "Design validation commands for docs-only, interface, service, logic, and runtime changes."
  • "Organize these tests into a CI gate and test matrix."

© hashgraph-online, Apache-2.0. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

SKILL.md and 2 other files (references) in plugins/yfge/agent-harness-skills/skills/validation-harness-design of hashgraph-online/awesome-codex-plugins.

  • SKILL.md
  • references/build-when-missing.md
  • references/environment-bootstrap.md

Open the folder on GitHubat commit 78497e5

Compare with similar skills

Validation Harness Design next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Validation Harness Design compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Validation Harness Design this skillhashgraph-online/awesome-codex-plugins1.2k—~866Automated safety check: PassApache-2.0
Debug Playwright Prowquay/quay2.8k—~2.2kAutomated safety check: PassApache-2.0
Quay Prow Triagequay/quay2.8k—~2.9kAutomated safety check: PassApache-2.0
Shopware CLIshopware/shopware-cli123—~4.9kAutomated safety check: PassMIT
Debug Surefireeclipse-rdf4j/rdf4j420—~2.4kAutomated safety check: PassBSD-3-Clause
Debug E2E Pipelinekubernetes-sigs/cloud-provider-azure294—~3.4kAutomated safety check: PassApache-2.0

Similar skills

  • Deep-dive diagnosis of a Playwright test failure already isolated to one Quay Prow/OpenShift CI run: downloads its GCS artifacts (results.json, JUnit, build/pod logs, Jaeger traces), classifies real…

    2.8k GitHub stars~2.2k tokensUpdated today
    Testing & QAAuto-check passed
  • Diagnose any Quay Prow job failure end to end: prowjob.json - top-level build log - JUnit - resolved failing step - Playwright results.json when the failing step is Playwright, continuing through…

    2.8k GitHub stars~2.9k tokensUpdated today
    Testing & QAAuto-check passed
  • Shopware CLI

    shopware/shopware-cli

    Use Shopware CLI for Shopware project, extension, and account workflows — create and install new projects (project create, project dev install), validate projects or extensions (with…

    123 GitHub stars~4.9k tokensUpdated yesterday
    Testing & QAAuto-check passed
  • Debug Surefire

    eclipse-rdf4j/rdf4j

    Debug Maven Surefire unit tests by running them in JDWP "wait for debugger" mode (-Dmaven.surefire.debug) and attaching to the forked test JVM using jdb (preferred for CLI/agent debugging)…

    420 GitHub stars~2.4k tokensUpdated yesterday
    Testing & QAAuto-check passed
  • Debug E2E Pipeline

    kubernetes-sigs/cloud-provider-azure

    Official

    Fetch and analyze Prow e2e pipeline failures for cloud-provider-azure.

    294 GitHub stars~3.4k tokensUpdated today
    Testing & QAAuto-check passed
  • Minecraft Testing

    Jahrome907/minecraft-agent-skills

    Design and implement automated tests for current Minecraft 26.x or legacy 1.21.x mods and plugins using JUnit, MockBukkit, NeoForge Game Tests, or Fabric Game Tests.

    166 GitHub stars~3.8k tokensUpdated 25 days ago
    Testing & QAAuto-check passed

More from hashgraph-online/awesome-codex-plugins

All 686 skills in this repo
  • Anime Reaction Gif

    hashgraph-online/awesome-codex-plugins

    Create original anime-style reaction stickers as looping GIFs and MP4 previews, using generated character pose sheets and timed key poses.

    1.2k GitHub stars~922 tokensUpdated today
    Auto-check passed
  • Calibredb

    hashgraph-online/awesome-codex-plugins

    Manage and query Calibre libraries with the calibredb CLI (local paths or Calibre Content server URLs).

    1.2k GitHub stars~1k tokensUpdated today
    Auto-check passed
  • Rust API Test Harness

    hashgraph-online/awesome-codex-plugins

    A skill your agent uses when adding, changing, testing, or debugging Rust HTTP APIs and services, especially when Codex needs black-box integration tests, random-port app startup, real database test…

    1.2k GitHub stars~1.7k tokensUpdated today
    Auto-check passed
  • Art

    hashgraph-online/awesome-codex-plugins

    Make a studio's game look like something at build time — a cover from a real frame of the game (free), painted covers, backdrops, textures and character plates from image models through the…

    1.2k GitHub stars~2.4k tokensUpdated today
    Auto-check passed
  • Game Balance Economy

    hashgraph-online/awesome-codex-plugins

    Balance game difficulty, resources, rewards, probability, progression, economies, and dominant strategies.

    1.2k GitHub stars~618 tokensUpdated today
    Auto-check passed
  • Manuscript Engagement Analytics

    hashgraph-online/awesome-codex-plugins

    Analyze nonfiction manuscripts for reader engagement signals, including heading-level word counts, slow starts, long slogs, weak takeaway titles, value pacing, beta-reader comment dropoff, and…

    1.2k GitHub stars~875 tokensUpdated today
    Auto-check passed

Works with

Categories

Questions about Validation Harness Design

What does Validation Harness Design do?

A skill your agent uses when designing repository validation commands, doctor scripts, test matrices, JSON or JUnit outputs, CI gates, smoke checks, or harness command surfaces. Validation Harness Design is an agent skill from hashgraph-online/awesome-codex-plugins. Use when designing repository validation commands, doctor scripts, test matrices, JSON or JUnit outputs, CI gates, smoke checks, or harness command surfaces.

When should I use Validation Harness Design?

Validation Harness Design fits situations like: designing repository validation commands; harness command surfaces.

How do I install Validation Harness Design in Claude Code?

Run `npx skills add hashgraph-online/awesome-codex-plugins --skill validation-harness-design -a claude-code`. Or copy the skill folder (plugins/yfge/agent-harness-skills/skills/validation-harness-design in hashgraph-online/awesome-codex-plugins) into .claude/skills/validation-harness-design in your project. Claude Code loads it when a task matches its description.

How do I install Validation Harness Design in Codex?

Run `npx skills add hashgraph-online/awesome-codex-plugins --skill validation-harness-design -a codex`. Or copy the skill folder (plugins/yfge/agent-harness-skills/skills/validation-harness-design in hashgraph-online/awesome-codex-plugins) into .agents/skills/validation-harness-design in your project. Codex loads it when a task matches its description.

Can I use Validation Harness Design in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add hashgraph-online/awesome-codex-plugins --skill validation-harness-design -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/validation-harness-design, .gemini/skills/validation-harness-design, .github/skills/validation-harness-design and .opencode/skills/validation-harness-design in your project.

What does Validation Harness Design need to run?

SKILL.md names no scripts, command-line tools or credentials: Validation Harness Design is instructions for the agent only.

Does Validation Harness Design access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is Validation Harness Design safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Validation Harness Design use?

Validation Harness Design is published under the Apache-2.0 licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Validation Harness Design use?

About 866 tokens (SKILL.md is roughly 3.5k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full. Its references folder adds about 778 tokens, read only when the agent opens those files.

What are the alternatives to Validation Harness Design?

Skills that share tags, products or a category with Validation Harness Design: Debug Playwright Prow (quay/quay, 2.8k stars), Quay Prow Triage (quay/quay, 2.8k stars), Shopware CLI (shopware/shopware-cli, 123 stars) and Debug Surefire (eclipse-rdf4j/rdf4j, 420 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Validation Harness Design?

hashgraph-online (a GitHub organization) maintains it in hashgraph-online/awesome-codex-plugins, which has 1,242 GitHub stars. The repository holds 686 skills in this directory. The repository was last updated on October 8, 2026.

Source: hashgraph-online/awesome-codex-plugins on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.