Agent skill

Exploratory Test

by tobihagemann in tobihagemann/turbo

Execute multi-level exploratory testing of the app covering basic functionality, complex operations, adversarial testing, and cross-cutting scenarios, plus usability observations through a UX lens…

MITAuto-check passedTesting & QA

Install Exploratory Test

skills CLI
$ npx skills add tobihagemann/turbo --skill exploratory-test -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install tobihagemann/turbo exploratory-test --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/tobihagemann/turbo.git skills-src && mkdir -p .claude/skills && cp -r skills-src/claude/skills/exploratory-test .claude/skills/exploratory-test && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
exploratory-test
GitHub stars
409
Token cost
~2k tokens
SKILL.md length
939 words
Files
1
Skills in repo
81
Repo updated
First seen
Licence
MIT

At a glance

Execute multi-level exploratory testing of the app covering basic functionality, complex operations, adversarial testing, and cross-cutting scenarios, plus usability observations through a UX lens…

  • Works in 6 steps: Load or Create Test Plan → Determine Testing Approach → Run /user-experience Skill (When… → …
  • The user asks to exploratory test
  • SKILL.md covers Task Tracking, Step 1: Load or Create Test Plan, Step 2: Determine Testing… and Step 3: Run /user-experience…, plus 4 more sections
  • Instructions only: no scripts, shell commands, URLs or credentials in SKILL.md

What it does

Exploratory Test is an agent skill from tobihagemann/turbo. Execute multi-level exploratory testing of the app covering basic functionality, complex operations, adversarial testing, and cross-cutting scenarios, plus usability observations through a UX lens reported separately from defects. Deeper than /smoke-test. Use when the user asks to "exploratory test", "test thoroughly", "test all scenarios", "deep test", "test edge cases", "test everything", "break it", "find bugs by testing", "test usability", or "check the UX while testing".

Its SKILL.md is about 2k tokens, which your agent loads only when the skill is triggered. It is a single SKILL.md file with no bundled scripts.

It sits in Testing & QA, covering QA and bug reports, UX design and Test generation. The repository describes itself as: Reusable workflows for planning, building, reviewing, and shipping with Claude Code and Codex. The licence is MIT.

When your agent uses it

  • The user asks to exploratory test
  • Test thoroughly
  • Test all scenarios
  • Test edge cases

Example prompts

  • “exploratory test”
  • “test thoroughly”
  • “test all scenarios”
  • “/exploratory-test”

Workflow steps

6 steps, taken from the step headings in SKILL.md.

  1. Load or Create Test Plan
  2. Determine Testing Approach
  3. Run /user-experience Skill (When User-Facing)
  4. Run /test-run-rules Skill
  5. Execute Tests by Level
  6. Report

What it can do on your machine

Read from SKILL.md and the folder at commit 160a0fa. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    No scripts in the folder and no shell commands in SKILL.md.

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Exploratory Test loads about 2k tokens when it runs. Until then it costs about 124 tokens; SKILL.md has 939 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~124
When it runs · the whole SKILL.md, loaded when a task matches
~2k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from tobihagemann/turbo at commit 160a0fa, republished under its MIT licence (© tobihagemann). 939 words, ~1,961 tokens.

Download SKILL.mdSave it as .claude/skills/exploratory-test/SKILL.md (or your agent's skills folder).
name
exploratory-test
description
Execute multi-level exploratory testing of the app covering basic functionality, complex operations, adversarial testing, and cross-cutting scenarios, plus usability observations through a UX lens reported separately from defects. Deeper than /smoke-test. Use when the user asks to "exploratory test", "test thoroughly", "test all scenarios", "deep test", "test edge cases", "test everything", "break it", "find bugs by testing", "test usability", or "check the UX while testing".

Exploratory Test

Execute multi-level exploratory testing that goes beyond smoke testing to actively find bugs through escalating test scenarios.

Task Tracking

At the start, use TaskCreate to create a task for each step:

  1. Load or create test plan
  2. Determine testing approach
  3. Run /user-experience skill (when user-facing)
  4. Run /test-run-rules skill
  5. Execute tests by level
  6. Report

Step 1: Load or Create Test Plan

Resolve the test plan using these rules in order:

  1. Explicit path — If a file path was passed, use it
  2. Explicit slug — resolve to .turbo/test-plans/<slug>.md
  3. Anchoring artifact — If the work under test is anchored to a plan, resolve to .turbo/test-plans/<that-slug>.md when that file exists
  4. Single file — Glob .turbo/test-plans/*.md. If exactly one file exists, use it
  5. Most recent — If multiple files exist, use the most recently modified
  6. Legacy fallback — .turbo/test-plan.md if .turbo/test-plans/ does not exist
  7. Nothing found — run the /create-test-plan skill first, then use the plan it writes

If multiple test plans exist and the most-recent choice is non-obvious, use AskUserQuestion to let the user pick from the candidates.

Read the resolved test plan and state its path.

Unless an explicit path or slug was passed, confirm the resolved plan still describes the work under test:

  • Unavailable branch state — a scenario's steps require a branch that no longer resolves in the repository
  • Completed prior run — every checkbox is already ticked and no recorded result is FAIL or PARTIAL
  • Superseded context — the plan's Context section names work that changes merged since the plan was written have reversed or removed

When a signal fires, output the signal and the scenarios it affects as text. For a superseded Context, name the scenarios that exercise the reversed or removed work. Then use AskUserQuestion to offer:

  • Regenerate — run the /create-test-plan skill with the resolved path, and use the plan it writes
  • Execute anyway — the signal is a false positive
  • Pick another plan — resolve to a different test plan file, then confirm that plan against these same signals

If the user specifies a narrower scope, filter the plan to relevant scenarios rather than executing all of them. Reserve filtering for that case: a superseded plan keeps scenarios that each look plausible alone, so trimming it preserves the wrong ones.

Step 2: Determine Testing Approach

Use the approach specified in the test plan. If the plan does not specify one, determine it using the same logic as /create-test-plan Step 2.

Step 3: Run /user-experience Skill (When User-Facing)

If the app has a user-facing surface (UI, screens, commands, messages, or any behavior a user sees or does), run the /user-experience skill to load the UX lens before executing tests, so usability concerns surface while interacting with the app. When it is unclear whether the surface is user-facing, use AskUserQuestion to ask rather than skipping silently. Skip this step for test targets with no user-facing behavior (internal library or infrastructure).

Step 4: Run /test-run-rules Skill

Run the /test-run-rules skill to load the rules for launching and driving the app.

Step 5: Execute Tests by Level

Work through each level sequentially. Complete all tests in a level before moving to the next.

Show full SKILL.md (415 more words)Show less
Execution Loop (Per Test)
  1. Set up the preconditions described in the test scenario
  2. Perform the exact steps
  3. Capture the result (screenshot, output, or state observation)
  4. Compare against the expected outcome
  5. Record PASS, FAIL, or PARTIAL with details
  6. When the UX lens is loaded, note any usability observation it surfaces, kept separate from the verdict

Record a scenario the test run rules leave blocked or inconclusive as PARTIAL, naming what is unproven and why.

When the scenario's output is consumed by another system, withhold PASS until that system accepts it. Decoding a token, reading a response body, or confirming a row exists shows only that the artifact was produced. Stand up the consumer under the same isolation and cleanup rules as any other service this run starts, and exercise its own flow. When standing it up is not possible, record PARTIAL and name which half is unproven. PARTIAL counts as not passed everywhere a verdict is tallied or gated.

Level Progression
  1. Level 1: Basic Functionality — If any Level 1 test does not pass, report early and use AskUserQuestion to ask whether to continue. Basic failures may indicate the feature is too broken for deeper testing.
  2. Level 2: Complex Operations — Execute all tests regardless of individual failures.
  3. Level 3: Adversarial Testing — Execute all tests. Failures here are expected and valuable.
  4. Level 4: Cross-Cutting Scenarios — Execute all tests.

If a project-specific testing skill or MCP tool was identified in Step 2, use that. The paths below are fallbacks.

Web App Path

Start or reuse a dev server under the test run rules. If /agent-browser is available, run the /agent-browser skill. Otherwise, use claude-in-chrome MCP to interact with the app.

UI/Native App Path

Launch the app. Use computer-use MCP to interact with the UI.

CLI Path

Run commands directly.

Step 6: Report

Present results organized by level:

Exploratory Test Results:

## Level 1: Basic Functionality (X/Y passed)
- [PASS] Test name: description — [substitution, when one was driven]
- [FAIL] Test name: description — [what went wrong]
- [PARTIAL] Test name: description — [what is unproven, and why]

## Level 2: Complex Operations (X/Y passed)
- [PASS] Test name: description — [substitution, when one was driven]
- [FAIL] Test name: description — [what went wrong]
- [PARTIAL] Test name: description — [what is unproven, and why]

## Level 3: Adversarial Testing (X/Y passed)
- [PASS] Test name: description — [substitution, when one was driven]
- [FAIL] Test name: description — [what went wrong]
- [PARTIAL] Test name: description — [what is unproven, and why]

## Level 4: Cross-Cutting Scenarios (X/Y passed)
- [PASS] Test name: description — [substitution, when one was driven]
- [FAIL] Test name: description — [what went wrong]
- [PARTIAL] Test name: description — [what is unproven, and why]

Overall: X/Y passed across all levels

Report usability observations from the UX lens below the level results, separately from the defects. A scenario can pass every functional check and still surface a usability concern.

## Usability Observations
- [UX] <observation> — names the UX context it touches (Understanding, Bridging, or Flowing) and the goal mismatch or friction it creates

For each failure, include the relevant screenshot, output, or state observation.

When the change under test spans several repositories, add a per-repo view of the findings below the usability observations, naming a suggested fix site for each.

Update the resolved test plan file by checking off completed tests and annotating results.

Then use the TaskList tool and proceed to any remaining task.

Rules

  • To diagnose failures, run the /investigate skill on the test report.

© tobihagemann, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

Just SKILL.md in claude/skills/exploratory-test of tobihagemann/turbo.

Open the folder on GitHubat commit 160a0fa

Compare with similar skills

Exploratory Test next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Exploratory Test compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Exploratory Test this skilltobihagemann/turbo409—~2kAutomated safety check: PassMIT
Scoutqa Testgithub/awesome-copilot40k1 repos~3.5kAutomated safety check: PassMIT
Bug Reproduction Test GeneratorArabelaTso/Skills-4-SE253—~1.8kAutomated safety check: PassApache-2.0
Test And Breakrohunj/claude-build-workflow231—~1.6kAutomated safety check: PassNone
Replica TestJakeschincariol/replica-skill1.2k—~819Automated safety check: PassMIT
Actionbook Web Testactionbook/actionbook1.6k—~9.7kAutomated safety check: PassApache-2.0

Similar skills

  • Scoutqa Test

    github/awesome-copilot

    Official

    This skill should be used when the user asks to "test this website", "run exploratory testing", "check for accessibility issues", "verify the login flow works", "find bugs on this page", or requests…

    40k GitHub starsUsed in 1 repo~3.5k tokens
    Testing & QAAuto-check passed
  • Bug Reproduction Test Generator

    ArabelaTso/Skills-4-SE

    Automatically generates executable tests that reproduce reported bugs from issue reports and code repositories.

    253 GitHub stars~1.8k tokensUpdated 1 mo ago
    Testing & QAAuto-check passed
  • Test And Break

    rohunj/claude-build-workflow

    Autonomous testing skill that opens a deployed app, goes through user flows, tries to break things, and writes detailed bug reports.

    231 GitHub stars~1.6k tokensUpdated 8 mo ago
    Testing & QAAuto-check passed
  • Replica Test

    Jakeschincariol/replica-skill

    Clicks through every flow of an app clone and tests it for bugs: a test plan generated from the recon flows with happy paths and edge cases, Playwright end-to-end tests where possible, a browser…

    1.2k GitHub stars~819 tokensUpdated 5 days ago
    Testing & QAAuto-check passed
  • Actionbook Web Test

    actionbook/actionbook

    Run browser-based web tests against websites using Actionbook CLI.

    1.6k GitHub stars~9.7k tokensUpdated 1 mo ago
    Testing & QAAuto-check passed
  • Test Plan

    quran/quran.com-frontend-next

    Generates a comprehensive testing plan based on the current branch changes or a specific PR.

    1.9k GitHub stars~1.4k tokensUpdated 5 mo ago
    Testing & QAAuto-check passed

More from tobihagemann/turbo

All 81 skills in this repo
  • Consult Oracle

    tobihagemann/turbo

    Consult ChatGPT Pro via ChatGPT browser automation for problems that resist standard approaches.

    409 GitHub stars~1.1k tokensUpdated today
    Auto-check passed
  • Fetch PR Comments

    tobihagemann/turbo

    Fetch and summarize review feedback and conversation from a GitHub PR (unresolved review threads, review bodies, and PR conversation comments) without making changes.

    409 GitHub stars~967 tokensUpdated today
    Auto-check passed
  • Recall Rationale

    tobihagemann/turbo

    Recall why a past change was made by locating the Claude Code transcript that produced it.

    409 GitHub stars~1.3k tokensUpdated today
    Auto-check passed
  • Resolve PR Comments

    tobihagemann/turbo

    Evaluate, fix, answer, and reply to GitHub pull request review comments and conversation comments.

    409 GitHub stars~3.8k tokensUpdated today
    Auto-check passed
  • Resolve PR Comments

    tobihagemann/turbo

    Evaluate, fix, answer, and reply to GitHub pull request review comments and conversation comments.

    409 GitHub stars~3.8k tokensUpdated today
    Auto-check passed
  • Assess Technical Debt

    tobihagemann/turbo

    Assess project-wide structural technical debt: complexity hotspots, deprecated API usage, duplication clusters, architecture rot, and low-value tests.

    409 GitHub stars~2.8k tokensUpdated today
    Auto-check passed

Categories

Questions about Exploratory Test

What does Exploratory Test do?

Execute multi-level exploratory testing of the app covering basic functionality, complex operations, adversarial testing, and cross-cutting scenarios, plus usability observations through a UX lens…. Exploratory Test is an agent skill from tobihagemann/turbo. Execute multi-level exploratory testing of the app covering basic functionality, complex operations, adversarial testing, and cross-cutting scenarios, plus usability observations through a UX lens reported separately from defects.

When should I use Exploratory Test?

Exploratory Test fits situations like: the user asks to exploratory test; test thoroughly; test all scenarios; test edge cases.

How do I install Exploratory Test in Claude Code?

Run `npx skills add tobihagemann/turbo --skill exploratory-test -a claude-code`. Or copy the skill folder (claude/skills/exploratory-test in tobihagemann/turbo) into .claude/skills/exploratory-test in your project. Claude Code loads it when a task matches its description.

How do I install Exploratory Test in Codex?

Run `npx skills add tobihagemann/turbo --skill exploratory-test -a codex`. Or copy the skill folder (claude/skills/exploratory-test in tobihagemann/turbo) into .agents/skills/exploratory-test in your project. Codex loads it when a task matches its description.

Can I use Exploratory Test in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add tobihagemann/turbo --skill exploratory-test -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/exploratory-test, .gemini/skills/exploratory-test, .github/skills/exploratory-test and .opencode/skills/exploratory-test in your project.

What does Exploratory Test need to run?

SKILL.md names no scripts, command-line tools or credentials: Exploratory Test is instructions for the agent only.

Does Exploratory Test access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is Exploratory Test safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Exploratory Test use?

Exploratory Test is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Exploratory Test use?

About 2k tokens (SKILL.md is roughly 7.8k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Exploratory Test?

Skills that share tags, products or a category with Exploratory Test: Scoutqa Test (github/awesome-copilot, 40k stars), Bug Reproduction Test Generator (ArabelaTso/Skills-4-SE, 253 stars), Test And Break (rohunj/claude-build-workflow, 231 stars) and Replica Test (Jakeschincariol/replica-skill, 1.2k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Exploratory Test?

tobihagemann (a GitHub user) maintains it in tobihagemann/turbo, which has 409 GitHub stars. The repository holds 81 skills in this directory. The repository was last updated on October 9, 2026.

Source: tobihagemann/turbo on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.