Agent skill

Accurate Testing

by theexperiencecompany in theexperiencecompany/gaia

Write adversarial test cases that break production code instead of producing false confidence.

Custom licenceAuto-check passedTesting & QA

Install Accurate Testing

skills CLI
$ npx skills add theexperiencecompany/gaia --skill accurate-testing -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install theexperiencecompany/gaia accurate-testing --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/theexperiencecompany/gaia.git skills-src && mkdir -p .claude/skills && cp -r skills-src/.agents/skills/accurate-testing .claude/skills/accurate-testing && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
accurate-testing
GitHub stars
308
Token cost
~3.4k tokens
SKILL.md length
1,476 words
Files
5 (incl. references)
Skills in repo
17
Repo updated
First seen
Licence
Custom licence

At a glance

Write adversarial test cases that break production code instead of producing false confidence.

  • Works in 3 steps: Every test must be falsifiable. There… → Prove falsifiability — don't assume it.… → Write tests against the code's blind…
  • Planning tests for any codebase
  • SKILL.md covers Prime Directive: Tests Exist…, The Mutation Check (The Real…, The Deletion Test (Minimum Gate) and Core Workflow, plus 7 more sections
  • Instructions only: no scripts, shell commands, URLs or credentials in SKILL.md

What it does

Accurate Testing is an agent skill from theexperiencecompany/gaia. Write adversarial test cases that break production code instead of producing false confidence. Use when writing, reviewing, or planning tests for any codebase. Triggers on: "write tests", "add test coverage", "test this function", "create unit tests", "integration tests", "fix flaky tests", "improve test coverage", "review test quality", "are these tests good", "test plan", "break this code", "find edge cases", or any task involving pytest, vitest, jest, or other test frameworks. Enforces the prime directive — a…

Its SKILL.md is about 3.4k tokens, which your agent loads only when the skill is triggered. The skill folder holds 5 other files, including reference files (for example `references/anti-patterns.md`, `references/breaking-the-code.md` and `references/pytest-patterns.md`).

It sits in Testing & QA, covering Unit testing and Test generation. It works with Jest, pytest and Vitest. The repository describes itself as: Your proactive personal AI assistant & companion for daily productivity 🌎.

When your agent uses it

  • Planning tests for any codebase
  • Add test coverage
  • Test this function
  • Create unit tests

Example prompts

  • “write tests”
  • “add test coverage”
  • “test this function”
  • “/accurate-testing”

Requirements

  • Python 3

Workflow steps

3 steps, taken from the first numbered list in SKILL.md.

  1. Every test must be falsifiable. There must exist a plausible bug that turns it red.
  2. Prove falsifiability — don't assume it. Break the code and watch the test fail (see the
  3. Write tests against the code's blind spots, not its happy path. After building a feature,

What it can do on your machine

Read from SKILL.md and the folder at commit 26424cb. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    No scripts in the folder and no shell commands in SKILL.md (its code samples are python).

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Accurate Testing loads about 3.4k tokens when it runs, and up to ~13k if it reads all its reference files. Until then it costs about 186 tokens; SKILL.md has 1,476 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~186
When it runs · the whole SKILL.md, loaded when a task matches
~3.4k
With references · SKILL.md plus every file in references/, read only if the agent opens them
~13k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

Its licence (Custom licence) doesn't allow us to republish the file, so here is its outline and opening line. It has 1,476 words (~3,408 tokens).

“A test is not a certificate that the code works. It is an attack on the code. You are not the author defending the implementation — you are the adversary trying to make it produce a wrong answer, crash, corrupt…”

— opening of SKILL.md by theexperiencecompany, Custom licence
name
accurate-testing

Read the full SKILL.md on GitHub

Files

SKILL.md and 4 other files (references) in .agents/skills/accurate-testing of theexperiencecompany/gaia.

  • SKILL.md
  • references/anti-patterns.md
  • references/breaking-the-code.md
  • references/pytest-patterns.md
  • references/vitest-patterns.md

Open the folder on GitHubat commit 26424cb

Compare with similar skills

Accurate Testing next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Accurate Testing compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Accurate Testing this skilltheexperiencecompany/gaia308—~3.4kAutomated safety check: PassCustom licence
MoAI TDD Workflowmodu-ai/moai-adk1.2k—~3.1kAutomated safety check: PassApache-2.0
Test Patternsrohitg00/skillkit1.5k—~1.8kAutomated safety check: PassApache-2.0
Test Generatoralirezarezvani/claude-code-tresor777—~1.6kAutomated safety check: PassMIT
TDD Guidealirezarezvani/claude-skills28k—~3.4kAutomated safety check: PassMIT
TDD GuideLeoYeAI/openclaw-master-skills2.2k—~1.4kAutomated safety check: PassMIT

Similar skills

  • MoAI TDD Workflow

    modu-ai/moai-adk

    Drives test-first development through the RED, GREEN, REFACTOR cycle, with a config switch that selects between TDD and a DDD workflow for existing code.

    1.2k GitHub stars~3.1k tokensUpdated today
    Testing & QAAuto-check passed
  • Test Patterns

    rohitg00/skillkit

    Applies proven testing patterns — Arrange-Act-Assert (AAA), Given-When-Then, Test Data Builders, Object Mother, parameterized tests, fixtures, spies, and test doubles — to help write maintainable…

    1.5k GitHub stars~1.8k tokensUpdated 4 mo ago
    Testing & QAAuto-check passed
  • Test Generator

    alirezarezvani/claude-code-tresor

    Automatically suggest tests for new functions and components.

    777 GitHub stars~1.6k tokensUpdated 3 mo ago
    Testing & QAAuto-check passed
  • TDD Guide

    alirezarezvani/claude-skills

    Test-driven development skill for writing unit tests, generating test fixtures and mocks, analyzing coverage gaps, and guiding red-green-refactor workflows across Jest, Pytest, JUnit, Vitest, and…

    28k GitHub stars~3.4k tokensUpdated 1 mo ago
    Testing & QAAuto-check passed
  • TDD Guide

    LeoYeAI/openclaw-master-skills

    Test-driven development skill for writing unit tests, generating test fixtures and mocks, analyzing coverage gaps, and guiding red-green-refactor workflows across Jest, Pytest, JUnit, Vitest, and…

    2.2k GitHub stars~1.4k tokensUpdated 2 mo ago
    Testing & QAAuto-check passed
  • TDD Guide

    borghei/Claude-Skills

    Guide red-green-refactor TDD with test generation, coverage-gap analysis, and multi- framework support.

    891 GitHub stars~1.7k tokensUpdated 2 days ago
    Testing & QAAuto-check passed

More from theexperiencecompany/gaia

All 17 skills in this repo
  • Audit Website

    theexperiencecompany/gaia

    Audit websites for SEO, performance, security, technical, content, and 15 other issue cateories with 230+ rules using the squirrelscan CLI.

    308 GitHub starsUsed in 1 repo~4.7k tokens
    Auto-check passed
  • Create DOCX

    theexperiencecompany/gaia

    Generate an editable Microsoft Word (.docx) document: reports, letters, memos, structured docs.

    308 GitHub stars~848 tokensUpdated today
    Auto-check passed
  • Create PDF

    theexperiencecompany/gaia

    Generate a polished, printable PDF: reports, letters, invoices, resumes, one-pagers.

    308 GitHub stars~1.3k tokensUpdated today
    Auto-check passed
  • Create PPTX

    theexperiencecompany/gaia

    Generate a PowerPoint (.pptx) slide deck: pitch decks, reviews, summaries.

    308 GitHub stars~895 tokensUpdated today
    Auto-check passed
  • Create Spreadsheet

    theexperiencecompany/gaia

    Generate an Excel (.xlsx) workbook or a CSV file: tables, financial models, data exports, formatted reports with charts.

    308 GitHub stars~802 tokensUpdated today
    Auto-check passed
  • Motion

    theexperiencecompany/gaia

    Build React animations with Motion (Framer Motion) - gestures (drag, hover, tap), scroll effects, spring physics, layout animations, SVG.

    308 GitHub stars~6.7k tokensUpdated today
    Auto-check passed

Categories

Questions about Accurate Testing

What does Accurate Testing do?

Write adversarial test cases that break production code instead of producing false confidence. Accurate Testing is an agent skill from theexperiencecompany/gaia. Write adversarial test cases that break production code instead of producing false confidence.

When should I use Accurate Testing?

Accurate Testing fits situations like: planning tests for any codebase; add test coverage; test this function; create unit tests.

How do I install Accurate Testing in Claude Code?

Run `npx skills add theexperiencecompany/gaia --skill accurate-testing -a claude-code`. Or copy the skill folder (.agents/skills/accurate-testing in theexperiencecompany/gaia) into .claude/skills/accurate-testing in your project. Claude Code loads it when a task matches its description.

How do I install Accurate Testing in Codex?

Run `npx skills add theexperiencecompany/gaia --skill accurate-testing -a codex`. Or copy the skill folder (.agents/skills/accurate-testing in theexperiencecompany/gaia) into .agents/skills/accurate-testing in your project. Codex loads it when a task matches its description.

Can I use Accurate Testing in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add theexperiencecompany/gaia --skill accurate-testing -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/accurate-testing, .gemini/skills/accurate-testing, .github/skills/accurate-testing and .opencode/skills/accurate-testing in your project.

What does Accurate Testing need to run?

SKILL.md names no scripts, command-line tools or credentials: Accurate Testing is instructions for the agent only. Our summary lists: Python 3.

Does Accurate Testing access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is Accurate Testing safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Accurate Testing use?

Accurate Testing has a licence file (the repository's licence) that doesn't match a standard licence. Read it on GitHub before reusing the skill.

How many tokens does Accurate Testing use?

About 3.4k tokens (SKILL.md is roughly 14k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full. Its references folder adds about 9.7k tokens, read only when the agent opens those files.

What are the alternatives to Accurate Testing?

Skills that share tags, products or a category with Accurate Testing: MoAI TDD Workflow (modu-ai/moai-adk, 1.2k stars), Test Patterns (rohitg00/skillkit, 1.5k stars), Test Generator (alirezarezvani/claude-code-tresor, 777 stars) and TDD Guide (alirezarezvani/claude-skills, 28k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Accurate Testing?

theexperiencecompany (a GitHub organization) maintains it in theexperiencecompany/gaia, which has 308 GitHub stars. The repository holds 17 skills in this directory. The repository was last updated on October 10, 2026.

Source: theexperiencecompany/gaia on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.