Agent skill

Testing

by rsmdt in rsmdt/the-startup

Writing effective tests and running them successfully. An agent skill from rsmdt/the-startup.

MITAuto-check passedTesting & QA

Install Testing

skills CLI
$ npx skills add rsmdt/the-startup --skill testing -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install rsmdt/the-startup testing --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/rsmdt/the-startup.git skills-src && mkdir -p .claude/skills && cp -r skills-src/plugins/team/skills/development/testing .claude/skills/testing && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
testing
GitHub stars
536
Token cost
~1.2k tokens
SKILL.md length
565 words
Files
2
Skills in repo
27
Repo updated
First seen
Licence
MIT

At a glance

Writing effective tests and running them successfully. An agent skill from rsmdt/the-startup.

  • Works in 5 steps: Assess Scope → Select Layer → Write Tests → …
  • Reviewing test quality
  • SKILL.md covers Persona, Interface, Constraints and Reference Materials, plus 1 more section
  • Instructions only: no scripts, shell commands, URLs or credentials in SKILL.md

What it does

Testing is an agent skill from rsmdt/the-startup. Writing effective tests and running them successfully. Covers layer-specific mocking rules, test design principles, debugging failures, and flaky test management. Use when writing tests, reviewing test quality, or debugging test failures.

Its SKILL.md is about 1.2k tokens, which your agent loads only when the skill is triggered. The skill folder holds 2 other files (for example `examples/test-pyramid.md`).

It sits in Testing & QA, covering Failing and flaky tests and Debugging. The repository describes itself as: The Agentic Startup - A collection of Claude Code commands, skills, and agents. The licence is MIT.

When your agent uses it

  • Reviewing test quality
  • Debugging test failures

Example prompts

  • “/testing”

Workflow steps

5 steps, taken from the step headings in SKILL.md.

  1. Assess Scope
  2. Select Layer
  3. Write Tests
  4. Run Tests
  5. Debug Failures

What it can do on your machine

Read from SKILL.md and the folder at commit 88d447c. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    No scripts in the folder and no shell commands in SKILL.md.

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Testing loads about 1.2k tokens when it runs. Until then it costs about 62 tokens; SKILL.md has 565 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~62
When it runs · the whole SKILL.md, loaded when a task matches
~1.2k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from rsmdt/the-startup at commit 88d447c, republished under its MIT licence (© rsmdt). 565 words, ~1,198 tokens.

Download SKILL.mdSave it as .claude/skills/testing/SKILL.md (or your agent's skills folder). This skill also uses 1 other file; get the full folder from GitHub.
name
testing
description
Writing effective tests and running them successfully. Covers layer-specific mocking rules, test design principles, debugging failures, and flaky test management. Use when writing tests, reviewing test quality, or debugging test failures.

Persona

Act as a testing specialist who writes effective tests, applies layer-appropriate mocking strategies, and debugs failures systematically. You enforce test quality standards and ensure the right behavior is tested at the right layer.

Test Context: $ARGUMENTS

Interface

TestDecision { layer: Unit | Integration | E2E mockingStrategy: string target: string pattern: ArrangeActAssert | GivenWhenThen }

DebugResult { failure: string rootCause: string fix: string }

State { context = $ARGUMENTS scope = null layer = null tests = [] failures = [] }

Constraints

Always:

  • Test behavior, not implementation — assert on observable outcomes.
  • One behavior per test — multiple assertions OK if verifying same logical outcome.
  • Use descriptive test names that state the expected behavior.
  • Follow Arrange-Act-Assert structure in every test.
  • Mock at boundaries only — databases, APIs, file system, time.
  • Use real internal collaborators — never mock application code.
  • Keep tests independent — no shared mutable state between tests.
  • Handle flaky tests aggressively — quarantine, fix within one week, or delete.
  • Focus on business-critical paths (payments, auth, core domain logic).
  • Prefer quality over quantity — 80% meaningful coverage beats 100% trivial coverage.

Never:

  • Mock internal methods or classes — that tests the mock, not the code.
  • Test implementation details — tests should survive refactoring.
  • Skip edge case testing — boundaries, null, empty, negative values.
  • Leave flaky tests in the main suite — they erode trust.

Reference Materials

Workflow

1. Assess Scope

Identify what needs testing:

match (context) { new feature code => write tests for new behavior bug fix => write regression test first, then fix refactoring => verify existing tests pass, add coverage gaps test review => evaluate test quality and coverage }

Determine layer distribution target:

  • Unit (60-70%) — isolated business logic
  • Integration (20-30%) — components with real dependencies
  • E2E (5-10%) — critical user journeys
2. Select Layer

match (scope) { business logic | validation | transformation | edge cases => Unit: mock at boundaries only, <100ms, no I/O, deterministic

database queries | API contracts | service communication | caching => Integration: real deps, mock external services only, <5s, clean state between tests

signup | checkout | auth flows | smoke tests => E2E: no mocking, real services in sandbox mode, <30s, critical paths only }

Mocking rules by layer:

  • Unit — mock external boundaries (DB, APIs, filesystem, time)
  • Integration — real databases, real caches, mock only third-party services
  • E2E — no mocking at all
Show full SKILL.md (213 more words)Show less
3. Write Tests

Apply Arrange-Act-Assert pattern. Name tests descriptively: "rejects order when inventory insufficient"

Always test edge cases:

  • Boundaries — min-1, min, min+1, max-1, max, max+1, zero, one, many
  • Special values — null, empty, negative, MAX_INT, NaN, unicode, leap years, timezones
  • Errors — network failures, timeouts, invalid input, unauthorized

Read examples/test-pyramid.md for layer-specific code examples.

4. Run Tests

Execute in order (fastest feedback first):

  1. Lint/typecheck
  2. Unit tests
  3. Integration tests
  4. E2E tests
5. Debug Failures

match (layer) { Unit => { 1. Read the assertion message carefully 2. Check test setup (Arrange section) 3. Run in isolation to rule out state leakage 4. Add logging to trace execution path } Integration => { 1. Check database state before/after 2. Verify mocks configured correctly 3. Look for race conditions or timing issues 4. Check transaction/rollback behavior } E2E => { 1. Check screenshots/videos 2. Verify selectors still match the UI 3. Add explicit waits for async operations 4. Run locally with visible browser 5. Compare CI environment to local } }

Flaky test protocol:

  1. Quarantine — move to separate suite immediately
  2. Fix within 1 week — or delete
  3. Common causes: shared state, time-dependent logic, race conditions, non-deterministic ordering

Anti-patterns to flag:

  • Over-mocking — testing mocks instead of code
  • Implementation test — breaks on refactoring
  • Shared state — test order affects results
  • Test duplication — use parameterized tests instead

© rsmdt, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

SKILL.md and 1 other file in plugins/team/skills/development/testing of rsmdt/the-startup.

  • SKILL.md
  • examples/test-pyramid.md

Open the folder on GitHubat commit 88d447c

Compare with similar skills

Testing next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Testing compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Testing this skillrsmdt/the-startup536—~1.2kAutomated safety check: PassMIT
Offloadimbue-ai/offload125—~3.1kAutomated safety check: PassMIT
Fix BugMelbourneDeveloper/dart_node1131 repos~709Automated safety check: NotesNone
Test Debuggingruby-git/ruby-git1.8k—~1.5kAutomated safety check: PassMIT
Fixavibebuilder/claude-prime120—~1kAutomated safety check: PassMIT
Bug Reproduction Test GeneratorArabelaTso/Skills-4-SE253—~1.8kAutomated safety check: PassApache-2.0

Similar skills

  • Offload

    imbue-ai/offload

    Activate when you see offload.toml in a repo, offload referenced in build targets (justfile, Makefile, scripts), or when you need to run a large test suite in parallel.

    125 GitHub stars~3.1k tokensUpdated 12 days ago
    Testing & QAAuto-check passed
  • Fix Bug

    MelbourneDeveloper/dart_node

    Fix a bug using test-driven development. An agent skill from MelbourneDeveloper/dart_node.

    113 GitHub starsUsed in 1 repo~709 tokens
    Testing & QAAuto-check: notes
  • Test Debugging

    ruby-git/ruby-git

    Debugs failing or flaky tests and improves test coverage. An agent skill from ruby-git/ruby-git.

    1.8k GitHub stars~1.5k tokensUpdated 6 days ago
    Testing & QAAuto-check passed
  • Fix

    avibebuilder/claude-prime

    Fix bugs and broken behavior when there is enough evidence to act on a repair path.

    120 GitHub stars~1k tokensUpdated 4 mo ago
    Testing & QAAuto-check passed
  • Bug Reproduction Test Generator

    ArabelaTso/Skills-4-SE

    Automatically generates executable tests that reproduce reported bugs from issue reports and code repositories.

    253 GitHub stars~1.8k tokensUpdated 1 mo ago
    Testing & QAAuto-check passed
  • Pester Failure Analysis

    PowerShell/PowerShell

    Investigates failing Pester tests in PowerShell CI jobs by following a six-step workflow from pull request status to documented fix recommendations.

    56k GitHub stars~5.1k tokensUpdated yesterday
    Testing & QAAuto-check passed

More from rsmdt/the-startup

All 27 skills in this repo
  • Analyze

    rsmdt/the-startup

    Deep-dive codebase analysis that explains how things actually work — business rules, architecture patterns, auth flows, data models, integrations, and performance hotspots.

    536 GitHub stars~1.9k tokensUpdated 2 mo ago
    Auto-check passed
  • Implement

    rsmdt/the-startup

    Implementation entry point. An agent skill from rsmdt/the-startup.

    536 GitHub stars~1.3k tokensUpdated 2 mo ago
    Auto-check passed
  • Implement Factory

    rsmdt/the-startup

    Factory loop orchestrator for multi-feature or multi-component implementation manifests.

    536 GitHub stars~3.2k tokensUpdated 2 mo ago
    Auto-check passed
  • Agentic Patterns

    rsmdt/the-startup

    Context enrichment for agentic AI application development using LangChain, Vercel AI SDK, and assistant-ui.

    536 GitHub stars~501 tokensUpdated 2 mo ago
    Auto-check passed
  • API Contract Design

    rsmdt/the-startup

    REST and GraphQL API design patterns, OpenAPI/Swagger specifications, versioning strategies, and authentication patterns.

    536 GitHub stars~1.1k tokensUpdated 2 mo ago
    Auto-check passed
  • Architecture Selection

    rsmdt/the-startup

    System architecture patterns including monolith, microservices, event-driven, and serverless, with C4 modeling, scalability strategies, and technology selection criteria.

    536 GitHub stars~1.2k tokensUpdated 2 mo ago
    Auto-check passed

Categories

Questions about Testing

What does Testing do?

Writing effective tests and running them successfully. An agent skill from rsmdt/the-startup. Testing is an agent skill from rsmdt/the-startup. Writing effective tests and running them successfully.

When should I use Testing?

Testing fits situations like: reviewing test quality; debugging test failures.

How do I install Testing in Claude Code?

Run `npx skills add rsmdt/the-startup --skill testing -a claude-code`. Or copy the skill folder (plugins/team/skills/development/testing in rsmdt/the-startup) into .claude/skills/testing in your project. Claude Code loads it when a task matches its description.

How do I install Testing in Codex?

Run `npx skills add rsmdt/the-startup --skill testing -a codex`. Or copy the skill folder (plugins/team/skills/development/testing in rsmdt/the-startup) into .agents/skills/testing in your project. Codex loads it when a task matches its description.

Can I use Testing in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add rsmdt/the-startup --skill testing -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/testing, .gemini/skills/testing, .github/skills/testing and .opencode/skills/testing in your project.

What does Testing need to run?

SKILL.md names no scripts, command-line tools or credentials: Testing is instructions for the agent only.

Does Testing access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is Testing safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Testing use?

Testing is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Testing use?

About 1.2k tokens (SKILL.md is roughly 4.8k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Testing?

Skills that share tags, products or a category with Testing: Offload (imbue-ai/offload, 125 stars), Fix Bug (MelbourneDeveloper/dart_node, 113 stars), Test Debugging (ruby-git/ruby-git, 1.8k stars) and Fix (avibebuilder/claude-prime, 120 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Testing?

rsmdt (a GitHub user) maintains it in rsmdt/the-startup, which has 536 GitHub stars. The repository holds 27 skills in this directory. The repository was last updated on August 3, 2026.

Source: rsmdt/the-startup on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.