Agent skill

Test Suite Analysis

by prime-radiant-inc in prime-radiant-inc/greenfield

Layer 1 skill for extracting behavioral intelligence from test suites.

Apache-2.0Auto-check passedTesting & QA

Install Test Suite Analysis

skills CLI
$ npx skills add prime-radiant-inc/greenfield --skill test-suite-analysis -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install prime-radiant-inc/greenfield test-suite-analysis --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/prime-radiant-inc/greenfield.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/test-suite-analysis .claude/skills/test-suite-analysis && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
test-suite-analysis
GitHub stars
292
Token cost
~3.2k tokens
SKILL.md length
1,054 words
Files
1
Skills in repo
21
Repo updated
First seen
Licence
Apache-2.0

At a glance

Layer 1 skill for extracting behavioral intelligence from test suites.

  • Works in 8 steps: Test code is RAW -- test files reference… → Assertions are behavioral contracts --… → E2E tests first -- prioritize end-to-end… → …
  • Tasks that involve Test generation
  • SKILL.md covers When to Use This Mode, Why Tests Are High-Value…, Framework Detection and Strategy 1: Read Test Code, plus 4 more sections
  • Calls go, npx and python

What it does

Test Suite Analysis is an agent skill from prime-radiant-inc/greenfield. Layer 1 skill for extracting behavioral intelligence from test suites. Framework detection, test code reading strategy, test execution strategy, behavioral claim extraction with Given/When/Then mapping, e2e vs unit value classification. Loaded by the analyzer agent during Layer 1.

Its SKILL.md is about 3.2k tokens, which your agent loads only when the skill is triggered. It is a single SKILL.md file with no bundled scripts.

It sits in Testing & QA, covering Test generation and End-to-end testing. It works with Playwright, Cypress, Jest and pytest. The repository describes itself as: A Claude Code plugin that reverse-engineers clean behavioral specs, test vectors, and acceptance criteria from any codebase, producing a provenance trail so a fresh team can… The licence is Apache-2.0.

When your agent uses it

  • Tasks that involve Test generation
  • Tasks that involve End-to-end testing

Example prompts

  • “/test-suite-analysis”

Requirements

  • Python 3
  • Node.js

Workflow steps

8 steps, taken from the first numbered list in SKILL.md.

  1. Test code is RAW -- test files reference internal implementation details. All output goes to workspace/raw/.
  2. Assertions are behavioral contracts -- treat every assertion as a confirmed behavioral claim. The developer is stating that this behavior…
  3. E2E tests first -- prioritize end-to-end and integration tests over unit tests. They encode user-visible behavior.
  4. Do not test mocked behavior -- a test that asserts a mock was called proves nothing about the system's actual behavior. Skip mock-only…
  5. Failed tests are data -- a failing test documents both the expected behavior (from the assertion) and the actual behavior (from the…
  6. Strategy 1 always works -- reading test code requires no environment. Always perform Strategy 1. Strategy 2 is additive and optional.
  7. Cite as you go -- every behavioral claim gets an inline <!-- cite: --> comment immediately after the claim. Never defer citation to a…
  8. Map the full surface -- do not stop after the first interesting test file. Build a complete inventory before deep-diving into individual…

What it can do on your machine

Read from SKILL.md and the folder at commit 6e6d4b4. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Shell commands in SKILL.md call:

    • go
    • npx
    • python
    • bundle

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md. Its commands use npx, which can reach the network depending on how they are called.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Test Suite Analysis loads about 3.2k tokens when it runs. Until then it costs about 75 tokens; SKILL.md has 1,054 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~75
When it runs · the whole SKILL.md, loaded when a task matches
~3.2k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from prime-radiant-inc/greenfield at commit 6e6d4b4, republished under its Apache-2.0 licence (© prime-radiant-inc). 1,054 words, ~3,169 tokens.

Download SKILL.mdSave it as .claude/skills/test-suite-analysis/SKILL.md (or your agent's skills folder).
name
test-suite-analysis
description
Layer 1 skill for extracting behavioral intelligence from test suites. Framework detection, test code reading strategy, test execution strategy, behavioral claim extraction with Given/When/Then mapping, e2e vs unit value classification. Loaded by the analyzer agent during Layer 1.

Test Suite Analysis Methodology

Extract behavioral intelligence from test suites. Tests are executable specifications -- they encode what the system MUST do in a form that can be verified. A passing test is a confirmed behavioral contract.

When to Use This Mode

Test suite analysis activates when:

  • The target repository contains test files
  • The discovery inventory identifies test files in the project
  • Other modes discover test directories during analysis

This mode runs independently of all other intelligence sources. All output is RAW (test code references internal implementation details).

Why Tests Are High-Value Intelligence

Tests are the only source type that is simultaneously:

  • Behavioral -- they describe what the system does, not how it's built
  • Executable -- they can be run to confirm the behavior still holds
  • Specific -- they provide exact inputs, expected outputs, and edge cases
  • Maintained -- failing tests get fixed, so they track current behavior

A single end-to-end test is worth more than a page of documentation because the test is verified by CI on every commit.

Framework Detection

Identify the test framework(s) in use before analyzing test code. Different frameworks use different assertion styles, test organization, and execution models.

FrameworkLanguageDetection Signals
JestJavaScript/TypeScriptjest.config.*, describe( / it( / expect( in __tests__/ or *.test.*, @jest/globals imports
PlaywrightJavaScript/TypeScriptplaywright.config.*, @playwright/test imports, page.goto( / page.click(
CypressJavaScript/TypeScriptcypress.config.*, cypress/ directory, cy.visit( / cy.get(
pytestPythonconftest.py, pytest.ini / pyproject.toml with [tool.pytest], files named test_*.py / *_test.py, assert statements
Go testingGo*_test.go files, testing.T / testing.B parameters, go test in CI config
RSpecRuby.rspec, spec/ directory, spec_helper.rb, describe / it / expect blocks
JUnitJava/Kotlin@Test annotations, src/test/ directory, assertEquals / assertThat calls
XCTestSwift/Objective-CXCTestCase subclasses, func test*() methods, XCTAssert* calls
Catch2C++#include <catch2/catch.hpp>, TEST_CASE( / SECTION( / REQUIRE( macros
Detection Strategy
bash
# Check for test configuration files
ls -la jest.config.* playwright.config.* cypress.config.* .rspec pytest.ini 2>/dev/null

# Check pyproject.toml for pytest config
grep -l '\[tool\.pytest' pyproject.toml 2>/dev/null

# Find test directories
find . -maxdepth 3 -type d \( -name "__tests__" -o -name "test" -o -name "tests" -o -name "spec" -o -name "cypress" \) 2>/dev/null

# Find test files by naming convention
find . -maxdepth 4 -type f \( -name "*.test.*" -o -name "*.spec.*" -o -name "test_*" -o -name "*_test.*" \) 2>/dev/null | head -50

# Count test files per pattern
echo "Jest/Mocha-style:" && find . -name "*.test.*" -o -name "*.spec.*" 2>/dev/null | wc -l
echo "Python-style:" && find . -name "test_*.py" -o -name "*_test.py" 2>/dev/null | wc -l
echo "Go-style:" && find . -name "*_test.go" 2>/dev/null | wc -l
echo "JUnit-style:" && find . -path "*/src/test/*" -name "*.java" 2>/dev/null | wc -l

Write detection results to workspace/raw/test-evidence/test-inventory.md.

Strategy 1: Read Test Code

Read test files directly and extract behavioral claims. This strategy always works -- it requires no working environment, no dependencies, and no execution.

1.1 Test File Inventory
bash
# Build complete inventory of test files with metadata
find . -type f \( -name "*.test.*" -o -name "*.spec.*" -o -name "test_*" -o -name "*_test.*" -o -name "*_test.go" \) 2>/dev/null | while read f; do
  lines=$(wc -l < "$f")
  echo "$lines $f"
done | sort -rn
1.2 Assertion Extraction

For each test file, extract the assertions -- these are the behavioral contracts:

bash
# Jest/Mocha assertions
grep -n "expect\|assert\|should\|toBe\|toEqual\|toContain\|toThrow\|toHaveBeenCalled" "$TEST_FILE"

# pytest assertions
grep -n "assert \|assert_\|assertEqual\|assertRaises\|pytest.raises" "$TEST_FILE"

# Go test assertions
grep -n "t\.Error\|t\.Fatal\|t\.Log\|assert\.\|require\." "$TEST_FILE"

# RSpec assertions
grep -n "expect\|should\|is_expected\|eq(\|include(\|raise_error" "$TEST_FILE"
1.3 Given/When/Then Extraction

Transform test code into behavioral claims using Given/When/Then structure:

For each test case (it(, test(, func Test*, def test_*), extract:

  • Given (setup/preconditions): fixture creation, mock configuration, state initialization
  • When (action): the function call, API request, or user action being tested
  • Then (assertions): the expected outcomes encoded in assertions
markdown
## Test: "should reject expired tokens"

**Given:** A token with expiry date in the past
**When:** The token is validated via `checkPermissions()`
**Then:**
- Returns false
- Sets error to "TOKEN_EXPIRED"
- Does not call the downstream service

**Source:** `auth.test.ts:45-62`
**Confidence:** confirmed (test assertion is an explicit behavioral contract)
1.4 E2E vs Unit Value Classification

Not all tests carry equal behavioral intelligence value:

Test TypeDetection SignalsBehavioral Value
End-to-end (e2e)Browser automation, HTTP requests to running server, multi-service interactionHigh -- tests the system as a user experiences it
IntegrationDatabase connections, external service calls, multi-module interactionHigh -- tests behavioral contracts between components
FunctionalSingle module tested with real dependenciesMedium -- tests module-level behavioral contracts
UnitMocked dependencies, isolated function testsLower -- tests implementation contracts, not user-visible behavior
SnapshottoMatchSnapshot(), toMatchInlineSnapshot()Low -- captures output format, not behavioral intent

Focus extraction effort on e2e and integration tests first. Unit tests fill in details after the behavioral surface is mapped.

1.5 Edge Case Mining

Tests are the richest source of edge case documentation. Look for:

bash
# Boundary value tests
grep -n "boundary\|limit\|max\|min\|overflow\|underflow\|zero\|empty\|null\|undefined" "$TEST_FILE" -i

# Error condition tests
grep -n "error\|fail\|reject\|throw\|invalid\|unauthorized\|forbidden\|timeout" "$TEST_FILE" -i

# Concurrency tests
grep -n "concurrent\|parallel\|race\|deadlock\|lock\|mutex\|async\|await" "$TEST_FILE" -i

# Special character and encoding tests
grep -n "unicode\|utf\|encoding\|escape\|special\|whitespace" "$TEST_FILE" -i

Write behavioral claims to workspace/raw/test-evidence/behavioral-claims.md. Write e2e flow documentation to workspace/raw/test-evidence/e2e-flows.md. Write edge case documentation to workspace/raw/test-evidence/edge-cases.md.

Strategy 2: Run Test Suite

Execute the test suite and observe its behavior. This strategy requires a working environment with all dependencies installed. It produces higher-confidence claims but has higher setup cost.

2.1 Prerequisites

Before attempting test execution:

  • Verify the container has all dependencies installed
  • Check for required environment variables or config files
  • Look for test setup scripts (beforeAll, setUp, fixtures, factories)
  • Identify tests that require external services (databases, APIs)
2.2 Execute with Maximum Verbosity
bash
# Jest
npx jest --verbose --no-coverage 2>&1 | tee workspace/raw/test-evidence/run-output.txt

# pytest
python -m pytest -v --tb=long 2>&1 | tee workspace/raw/test-evidence/run-output.txt

# Go
go test -v ./... 2>&1 | tee workspace/raw/test-evidence/run-output.txt

# RSpec
bundle exec rspec --format documentation 2>&1 | tee workspace/raw/test-evidence/run-output.txt
2.3 Observe Runtime Behavior

During test execution, capture:

  • Network calls -- tests that make HTTP requests reveal API contracts
  • File I/O -- tests that read/write files reveal data format contracts
  • Timing -- slow tests may indicate external dependency interaction
  • Failures -- failed tests reveal behavioral regressions or environment-specific behavior
  • Warnings -- deprecation warnings and lint output reveal upcoming behavioral changes
bash
# Capture network activity during tests (if strace/dtrace available)
strace -e trace=network -f npx jest 2> workspace/raw/test-evidence/network-trace.txt

# Capture file I/O during tests
strace -e trace=file -f npx jest 2> workspace/raw/test-evidence/file-trace.txt
Show full SKILL.md (400 more words)Show less
2.4 Failure Analysis

Failed tests are behavioral intelligence:

  • A failing test documents a behavioral contract that is currently violated
  • The expected value in the assertion documents what the behavior SHOULD be
  • The actual value documents what the behavior currently IS
  • The gap between expected and actual is a behavioral specification

Record every failure with:

  • Test name and location
  • Expected behavior (from the assertion)
  • Actual behavior (from the error output)
  • Whether this appears to be a genuine regression or an environment issue

Provenance Rules

Source Type

All claims from test suite analysis use source=test-suite:

markdown
- Expired tokens are rejected with a TOKEN_EXPIRED error
  <!-- cite: source=test-suite, ref=src/auth/__tests__/auth.test.ts:45-62, confidence=confirmed, agent=test-analyzer -->
Confidence Levels
  • confirmed -- the test assertion explicitly encodes the behavioral claim AND the test passes (or the claim is directly readable from the test code regardless of execution)
  • inferred -- the behavioral claim is derived from test setup, fixture data, or mock configuration rather than direct assertion
  • assumed -- the behavioral claim is derived from test naming, file organization, or structural patterns rather than assertion content
Test Code Assertions Are Confirmed

Unlike source code analysis (where claims are typically inferred), behavioral claims extracted directly from test assertions are confirmed. A test assertion is an explicit, executable behavioral contract. The developer who wrote expect(result).toBe(42) is asserting that this behavior is required.

Cite As You Go

Every behavioral claim gets an inline citation immediately after the claim. The ref field should be <file-path>:<line-range>.

Output Structure

workspace/raw/test-evidence/
    test-inventory.md          # Framework detection, test file inventory, classification
    behavioral-claims.md       # Behavioral claims in Given/When/Then format
    e2e-flows.md               # End-to-end flow documentation from e2e/integration tests
    edge-cases.md              # Edge cases, boundary conditions, error handling from tests
    run-output.txt             # Raw test execution output (if Strategy 2 was used)

Rules

  1. Test code is RAW -- test files reference internal implementation details. All output goes to workspace/raw/.
  2. Assertions are behavioral contracts -- treat every assertion as a confirmed behavioral claim. The developer is stating that this behavior is required.
  3. E2E tests first -- prioritize end-to-end and integration tests over unit tests. They encode user-visible behavior.
  4. Do not test mocked behavior -- a test that asserts a mock was called proves nothing about the system's actual behavior. Skip mock-only tests when extracting behavioral claims.
  5. Failed tests are data -- a failing test documents both the expected behavior (from the assertion) and the actual behavior (from the error). Record both.
  6. Strategy 1 always works -- reading test code requires no environment. Always perform Strategy 1. Strategy 2 is additive and optional.
  7. Cite as you go -- every behavioral claim gets an inline <!-- cite: --> comment immediately after the claim. Never defer citation to a later step.
  8. Map the full surface -- do not stop after the first interesting test file. Build a complete inventory before deep-diving into individual files.

© prime-radiant-inc, Apache-2.0. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

Just SKILL.md in skills/test-suite-analysis of prime-radiant-inc/greenfield.

Open the folder on GitHubat commit 6e6d4b4

Compare with similar skills

Test Suite Analysis next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Test Suite Analysis compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Test Suite Analysis this skillprime-radiant-inc/greenfield292—~3.2kAutomated safety check: PassApache-2.0
Test CommanderEliasOulkadi/shokunin114—~3kAutomated safety check: NotesMIT
Test Detectdavila7/claude-code-templates33k—~989Automated safety check: PassMIT
Testing Patternssoftspark/ai-toolkit179—~1.6kAutomated safety check: PassApache-2.0
Error Explanation GeneratorArabelaTso/Skills-4-SE253—~3.8kAutomated safety check: PassApache-2.0
Angular TestingKilo-Org/kilo-marketplace190—~2.9kAutomated safety check: PassMIT

Similar skills

  • Test Commander

    EliasOulkadi/shokunin

    Generate unit, integration, E2E, and visual regression tests following the Testing Trophy methodology (80% integration).

    114 GitHub stars~3k tokensUpdated 6 days ago
    Testing & QAAuto-check: notes
  • Test Detect

    davila7/claude-code-templates

    Auto-detect testing framework and run relevant tests. An agent skill from davila7/claude-code-templates.

    33k GitHub stars~989 tokensUpdated yesterday
    Testing & QAAuto-check passed
  • Testing Patterns

    softspark/ai-toolkit

    Testing strategy: pyramid, AAA, mocks/fakes/stubs, flaky tests, coverage.

    179 GitHub stars~1.6k tokensUpdated 2 days ago
    Testing & QAAuto-check passed
  • Error Explanation Generator

    ArabelaTso/Skills-4-SE

    Explains test failures and provides actionable debugging guidance.

    253 GitHub stars~3.8k tokensUpdated 1 mo ago
    Testing & QAAuto-check passed
  • Angular Testing

    Kilo-Org/kilo-marketplace

    Write unit and integration tests for Angular v20+ applications using Vitest or Jasmine with TestBed and modern testing patterns.

    190 GitHub stars~2.9k tokensUpdated 12 days ago
    Testing & QAAuto-check passed
  • Testing Strategy Builder

    aiskillstore/marketplace

    A skill your agent uses when creating comprehensive testing strategies for applications.

    433 GitHub stars~3.6k tokensUpdated yesterday
    Testing & QAAuto-check passed

More from prime-radiant-inc/greenfield

All 21 skills in this repo
  • Reverse Engineering Analysis Pipeline

    prime-radiant-inc/greenfield

    Master methodology for reverse-engineering a codebase into behavioral specs with cited evidence, reading every line across source, binaries, docs, runtime and git history.

    292 GitHub stars~3.6k tokensUpdated 2 mo ago
    Auto-check passed
  • Community Intelligence Research

    prime-radiant-inc/greenfield

    Mines tutorials, forums, reviews, issues and changelogs for observed product behavior, using six search channels and consensus analysis.

    292 GitHub stars~4.5k tokensUpdated 2 mo ago
    Auto-check passed
  • Containerized Target Execution

    prime-radiant-inc/greenfield

    Runs untrusted analysis targets inside Docker or Podman containers with memory, CPU and process limits, covering image builds, lifecycle, command execution and cleanup.

    292 GitHub stars~2.1k tokensUpdated 2 mo ago
    Auto-check passed
  • API Contract Detection

    prime-radiant-inc/greenfield

    Finds OpenAPI, GraphQL, Protobuf and JSON Schema files in a codebase and extracts behavioral claims from them as part of a reverse-engineering workflow.

    292 GitHub stars~4.2k tokensUpdated 2 mo ago
    Auto-check passed
  • Documentation Research Methodology

    prime-radiant-inc/greenfield

    Method for extracting behavioral specifications from a product's public documentation: tiered search order, claim extraction rules, output structure, stop criteria and gap analysis.

    292 GitHub stars~4.6k tokensUpdated 2 mo ago
    Auto-check passed
  • Ecosystem Analysis

    prime-radiant-inc/greenfield

    Layer 1 skill for SDK and ecosystem analysis. An agent skill from prime-radiant-inc/greenfield.

    292 GitHub stars~2.9k tokensUpdated 2 mo ago
    Auto-check passed

Categories

Questions about Test Suite Analysis

What does Test Suite Analysis do?

Layer 1 skill for extracting behavioral intelligence from test suites. Test Suite Analysis is an agent skill from prime-radiant-inc/greenfield. Layer 1 skill for extracting behavioral intelligence from test suites.

When should I use Test Suite Analysis?

Test Suite Analysis fits situations like: tasks that involve Test generation; tasks that involve End-to-end testing.

How do I install Test Suite Analysis in Claude Code?

Run `npx skills add prime-radiant-inc/greenfield --skill test-suite-analysis -a claude-code`. Or copy the skill folder (skills/test-suite-analysis in prime-radiant-inc/greenfield) into .claude/skills/test-suite-analysis in your project. Claude Code loads it when a task matches its description.

How do I install Test Suite Analysis in Codex?

Run `npx skills add prime-radiant-inc/greenfield --skill test-suite-analysis -a codex`. Or copy the skill folder (skills/test-suite-analysis in prime-radiant-inc/greenfield) into .agents/skills/test-suite-analysis in your project. Codex loads it when a task matches its description.

Can I use Test Suite Analysis in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add prime-radiant-inc/greenfield --skill test-suite-analysis -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/test-suite-analysis, .gemini/skills/test-suite-analysis, .github/skills/test-suite-analysis and .opencode/skills/test-suite-analysis in your project.

What does Test Suite Analysis need to run?

Going by SKILL.md and its folder, Test Suite Analysis needs the command-line tools its instructions call (go, npx, python and bundle). Our summary lists: Python 3; Node.js.

Does Test Suite Analysis access the network?

SKILL.md contains no URLs. Its commands use npx, which can reach the network depending on how they are called. This is read from the text; nothing was executed.

Is Test Suite Analysis safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Test Suite Analysis use?

Test Suite Analysis is published under the Apache-2.0 licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Test Suite Analysis use?

About 3.2k tokens (SKILL.md is roughly 13k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Test Suite Analysis?

Skills that share tags, products or a category with Test Suite Analysis: Test Commander (EliasOulkadi/shokunin, 114 stars), Test Detect (davila7/claude-code-templates, 33k stars), Testing Patterns (softspark/ai-toolkit, 179 stars) and Error Explanation Generator (ArabelaTso/Skills-4-SE, 253 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Test Suite Analysis?

prime-radiant-inc (a GitHub organization) maintains it in prime-radiant-inc/greenfield, which has 292 GitHub stars. The repository holds 21 skills in this directory. The repository was last updated on August 6, 2026.

Source: prime-radiant-inc/greenfield on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.