Agent skill

Testing

by nteract in nteract/nteract

Run tests, verify changes, and collect diagnostics. An agent skill from nteract/nteract.

BSD-3-ClauseAuto-check passedTesting & QA

Install Testing

skills CLI
$ npx skills add nteract/nteract --skill testing -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install nteract/nteract testing --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/nteract/nteract.git skills-src && mkdir -p .claude/skills && cp -r skills-src/.agents/skills/testing .claude/skills/testing && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
testing
GitHub stars
179
Token cost
~2.2k tokens
SKILL.md length
667 words
Files
1
Skills in repo
10
Repo updated
First seen
Licence
BSD-3-Clause

At a glance

Run tests, verify changes, and collect diagnostics. An agent skill from nteract/nteract.

  • Writing new tests
  • SKILL.md covers Quick Reference, Verification Workflow, Frontend Unit Tests (Vitest) and Rust Unit Tests, plus 6 more sections
  • Calls cargo, pnpm and pytest
  • Verifying code changes before commit

What it does

Testing is an agent skill from nteract/nteract. Run tests, verify changes, and collect diagnostics. Use when running tests, writing new tests, verifying code changes before commit, or collecting logs for debugging.

Its SKILL.md is about 2.2k tokens, which your agent loads only when the skill is triggered. It is a single SKILL.md file with no bundled scripts.

It sits in Testing & QA, covering Unit testing and End-to-end testing. It works with Python, pnpm and pytest. The repository describes itself as: We're back! Now firing notebooks out of a t-shirt gun. The licence is BSD-3-Clause.

When your agent uses it

  • Writing new tests
  • Verifying code changes before commit
  • Collecting logs for debugging

Example prompts

  • “/testing”

Requirements

  • Python 3

What it can do on your machine

Read from SKILL.md and the folder at commit 414222e. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Shell commands in SKILL.md call:

    • cargo
    • pnpm
    • pytest
    • python
    • pip
    • deno

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md. Its commands use pnpm and pip, which can reach the network depending on how they are called.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Testing loads about 2.2k tokens when it runs. Until then it costs about 44 tokens; SKILL.md has 667 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~44
When it runs · the whole SKILL.md, loaded when a task matches
~2.2k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from nteract/nteract at commit 414222e, republished under its BSD-3-Clause licence (© nteract). 667 words, ~2,202 tokens.

Download SKILL.mdSave it as .claude/skills/testing/SKILL.md (or your agent's skills folder).
name
testing
description
Run tests, verify changes, and collect diagnostics. Use when running tests, writing new tests, verifying code changes before commit, or collecting logs for debugging.

Testing & Verification

Quick Reference

TypeLocationCommandFramework
Browser E2Eapps/notebook/e2e/pnpm --filter notebook-ui test:e2e:browserPlaywright
Native E2Ee2e/specs/cargo xtask e2e testWebdriverIO + Mocha
Frontend unitsrc/**/__tests__/, apps/notebook/src/**/__tests__/pnpm testVitest + jsdom
Rust unitinline #[cfg(test)]cargo testbuilt-in
CLI behaviorcrates/runt/tests/*.honecargo hone testHone
Pythonpython/runtimed/tests/pytestpytest

Verification Workflow

After making changes, run the narrowest credible test first, then broader checks.

Narrow Tests by Crate
Files changedTest command
crates/runtimed/src/**cargo test -p runtimed
crates/notebook-wire/src/**cargo test -p notebook-wire && cargo test -p notebook-protocol
crates/notebook-doc/src/**cargo test -p notebook-doc
crates/notebook-protocol/src/**cargo test -p notebook-protocol
crates/notebook-sync/src/**cargo test -p notebook-sync
crates/kernel-env/src/**cargo test -p kernel-env
crates/kernel-launch/src/**cargo test -p kernel-launch
crates/runt/src/**cargo test -p runt
crates/runt-workspace/src/**cargo test -p runt-workspace
crates/runtimed-py/src/**up rebuild=true
crates/runtimed-wasm/**cargo xtask wasm then deno test --allow-read --allow-env --no-check
apps/notebook/src/**pnpm test:run
python/runtimed/src/**pytest python/runtimed/tests/test_session_unit.py -v

Multiple crates: cargo test -p runtimed -p notebook-doc.

MCP Live Verification (when nteract-dev available)

For daemon/kernel changes: up rebuild=true → create_notebook → create_cell with 1 + 1 → execute_cell → verify output is 2.

For CRDT/doc changes: create_notebook → create_cell → get_cell (verify source) → set_cell → get_cell (verify update).

For kernel-env changes: up rebuild=true → create_notebook → execute import sys; print(sys.executable) → verify Python path.

Reporting verification

State which checks ran and what they established. Distinguish compilation, narrow tests, and MCP live verification; name skipped or blocked checks and the behavior they leave unverified.

For example, after narrow tests pass but a live check is unavailable:

The daemon tests passed. I couldn't run the MCP check because nteract-dev wasn't available, so live cell execution is still unverified.

Always run cargo xtask lint --fix before committing.

Frontend Unit Tests (Vitest)

Config: vitest.config.ts (jsdom environment, globals enabled).

bash
pnpm test         # Watch mode
pnpm test:run     # Run once

Key locations: src/components/isolated/__tests__/, src/components/outputs/__tests__/, src/components/widgets/__tests__/, apps/notebook/src/lib/__tests__/.

Rust Unit Tests

bash
cargo test                    # All workspace tests
cargo test -p runtimed        # Specific crate
cargo test -- --nocapture     # Show println! output

Hone CLI Tests

Declarative bash-based tests in crates/runt/tests/*.hone.

bash
cargo hone test               # All hone tests
cargo hone test cli.hone      # Specific file

Assertions: ASSERT exit_code == 0, ASSERT stdout contains "text", ASSERT stdout matches /pattern/.

Python Tests

Two venvs: workspace (.venv at root) for dev, and python/runtimed/.venv for isolated pytest.

bash
# Setup test venv
cd python/runtimed && python -m venv .venv && source .venv/bin/activate
pip install -e ".[dev]"
cd ../../crates/runtimed-py && VIRTUAL_ENV=../../python/runtimed/.venv maturin develop

# Run
pytest python/runtimed/tests/test_session_unit.py -v          # Unit (no daemon)
SKIP_INTEGRATION_TESTS=1 pytest python/runtimed/tests/ -v     # Skip integration
RUNTIMED_INTEGRATION_TEST=1 pytest python/runtimed/tests/ -v  # CI mode (spawns daemon)

Browser E2E Tests

Playwright tests in apps/notebook/e2e/ are the preferred E2E surface for notebook UI, runtime, fixture, and renderer behavior. They use the Vite browser relay plus the dev daemon, avoiding the slower Tauri WebDriver app build.

bash
pnpm --filter notebook-ui test:e2e:browser -- --project=chromium
pnpm --filter notebook-ui test:e2e:browser -- --project=chromium apps/notebook/e2e/execute-cell-output.spec.ts
pnpm --filter notebook-ui test:e2e:browser:portless -- --project=chromium

Native E2E Tests

Native WebDriverIO tests should stay focused on Tauri-specific setup: app shell boot, cwd-derived untitled notebook behavior, launcher/bootstrap behavior, and other issues that require the real Tauri process.

Running
bash
cargo xtask e2e build       # Build with WebDriver support (required first)
cargo xtask e2e test        # Native smoke/default run
cargo xtask e2e test-all    # Native smoke plus long-tail Tauri fixtures
cargo xtask e2e test-fixture <notebook> <spec>  # Single fixture test
Show full SKILL.md (279 more words)Show less
Adding Tests

Fixture test: Create notebook in crates/notebook/fixtures/audit-test/, create spec in e2e/specs/, add to FIXTURE_SPECS in e2e/wdio.conf.js, add to crates/xtask/src/main.rs, add to CI.

Regular test: Create a spec in e2e/specs/. It is picked up automatically if not in FIXTURE_SPECS.

Helpers (e2e/helpers.js)
HelperPurpose
waitForAppReady()Waits for toolbar (15s). Use in every before() hook
waitForKernelReady()Waits for kernel idle/busy (60s). Superset of above
executeFirstCell()Focuses first code cell, Shift+Enter
waitForCellOutput(cell)Waits for stream output
waitForOutputContaining(cell, text)Waits for specific output text
approveTrustDialog()Clicks "Trust & Install"
typeSlowly(text)Character-by-character (30ms). Required for CodeMirror
setupCodeCell()Finds/creates code cell, focuses editor, selects all
wry WebDriver Constraints
  • Use data-testid attributes. Text selectors return broken refs.
  • Use browser.execute() + browser.waitUntil(). executeAsync() is unsupported.
  • Use typeSlowly() for CodeMirror. Fast input drops characters.
  • Use browser.execute() for iframe testing. switchToFrame() is broken.
Selectors

data-testid: notebook-toolbar, save-button, add-code-cell-button, add-markdown-cell-button, start-kernel-button, restart-kernel-button, interrupt-kernel-button, run-all-button, deps-toggle, trust-dialog, trust-approve-button, deps-panel, deps-add-input.

data-slot: output-area, ansi-stream-output, ansi-error-output.

Other: [data-cell-type="code"], [data-cell-type="markdown"], .cm-content[contenteditable="true"], iframe[sandbox].

Diagnostics

Collecting

Use env -i for system diagnostics to avoid dev env vars (RUNTIMED_DEV, RUNTIMED_WORKSPACE_PATH) leaking through. env -i also clears PATH, so use the installed nteract link (~/.local/bin by default) and --channel to pick the installation.

bash
# Nightly (system)
env -i HOME=$HOME "$HOME/.local/bin/nteract" --channel nightly diagnostics

# Stable (system)
env -i HOME=$HOME "$HOME/.local/bin/nteract" --channel stable diagnostics

# Dev daemon (no env -i needed)
RUNTIMED_DEV=1 RUNTIMED_WORKSPACE_PATH="$(pwd)" ./target/debug/runt diagnostics

Other system commands follow the same env -i pattern:

bash
env -i HOME=$HOME "$HOME/.local/bin/nteract" --channel nightly daemon status
env -i HOME=$HOME "$HOME/.local/bin/nteract" --channel nightly daemon logs -f
env -i HOME=$HOME "$HOME/.local/bin/nteract" --channel stable ps

Read files from tarball without extracting:

bash
tar xzf <archive>.tar.gz -O doctor.json
tar xzf <archive>.tar.gz -O runtimed.log | grep -i 'error\|panic'
What to Look For
  • Ghost windows: Context for '...' missing in notebook.log
  • Daemon crashes: Check runtimed.log.1 (previous session)
  • Upgrade failures: Search [upgrade] in notebook.log
  • Kernel issues: Search [daemon-kernel] or kernel_status
  • Sync errors: Search [notebook-sync] or daemon:disconnected
  • Frontend errors: webview:error or webview:warn in notebook.log
  • launchd issues: Check doctor.json launchd_service status

Test Philosophy

Prefer fast integration tests over slow E2E. Use E2E for critical user journeys, integration tests for daemon behavior, unit tests for algorithms.

© nteract, BSD-3-Clause. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

Just SKILL.md in .agents/skills/testing of nteract/nteract.

Open the folder on GitHubat commit 414222e

Compare with similar skills

Testing next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Testing compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Testing this skillnteract/nteract179—~2.2kAutomated safety check: PassBSD-3-Clause
Python Testing Patternsjh941213/my-cc-harness12617 repos~5.4kAutomated safety check: PassNone
Specx Testsmaksimzayats/specx202—~1.9kAutomated safety check: PassMIT
Python Testingmacalbert/envilder138—~3.1kAutomated safety check: PassMIT
Testing Patternssoftspark/ai-toolkit179—~1.6kAutomated safety check: PassApache-2.0
Testing Pyericrisco/rsc-harness167—~3.2kAutomated safety check: PassMIT

Similar skills

  • Python Testing Patterns

    jh941213/my-cc-harness

    Implement comprehensive testing strategies with pytest, fixtures, mocking, and test-driven development.

    126 GitHub starsUsed in 17 repos~5.4k tokens
    Testing & QAAuto-check passed
  • Specx Tests

    maksimzayats/specx

    Add or refine tests for specx Python services. An agent skill from maksimzayats/specx.

    202 GitHub stars~1.9k tokensUpdated 2 mo ago
    Testing & QAAuto-check passed
  • Python Testing

    macalbert/envilder

    Mandatory testing conventions including AAA pattern, test naming, assertions, and mocks.

    138 GitHub stars~3.1k tokensUpdated 3 days ago
    Testing & QAAuto-check passed
  • Testing Patterns

    softspark/ai-toolkit

    Testing strategy: pyramid, AAA, mocks/fakes/stubs, flaky tests, coverage.

    179 GitHub stars~1.6k tokensUpdated yesterday
    Testing & QAAuto-check passed
  • Testing Py

    ericrisco/rsc-harness

    A skill your agent uses when writing or fixing Python tests with pytest — especially when the suite is green but bugs still ship, when you must decide where to patch a mocked dependency, when…

    167 GitHub stars~3.2k tokensUpdated today
    Testing & QAAuto-check passed
  • Error Explanation Generator

    ArabelaTso/Skills-4-SE

    Explains test failures and provides actionable debugging guidance.

    253 GitHub stars~3.8k tokensUpdated 1 mo ago
    Testing & QAAuto-check passed

More from nteract/nteract

All 10 skills in this repo
  • Automerge Sync

    nteract/nteract

    Automerge sync protocol internals, document model (OpSet, ChangeGraph, fork/merge, save/load lifecycle), and higher-level protocol design patterns.

    179 GitHub stars~4.5k tokensUpdated today
    Auto-check passed
  • Daemon Dev

    nteract/nteract

    Develop, debug, and manage the runtimed daemon, Python bindings, and build system.

    179 GitHub stars~3k tokensUpdated today
    Auto-check passed
  • Execution Pipeline

    nteract/nteract

    The end-to-end cell execution pipeline from MCP tool call through daemon to kernel and back.

    179 GitHub stars~2.7k tokensUpdated today
    Auto-check passed
  • Nteract Diagnostics

    nteract/nteract

    Pull and triage submitted nteract diagnostics archives from Cloudflare using a diagnostics id/token.

    179 GitHub stars~1.1k tokensUpdated today
    Auto-check passed
  • Repl

    nteract/nteract

    Use nteract notebooks as a persistent Python REPL. An agent skill from nteract/nteract.

    179 GitHub stars~1.1k tokensUpdated today
    Auto-check passed
  • Architecture

    nteract/nteract

    Architecture and documentation framing for cross-cutting repo decisions, docs taxonomy placement, ADRs, memos, PRDs, implementation plans, audits, measurements, runbooks, and source-grounded…

    179 GitHub stars~497 tokensUpdated today
    Auto-check passed

Categories

Questions about Testing

What does Testing do?

Run tests, verify changes, and collect diagnostics. An agent skill from nteract/nteract. Testing is an agent skill from nteract/nteract. Run tests, verify changes, and collect diagnostics.

When should I use Testing?

Testing fits situations like: writing new tests; verifying code changes before commit; collecting logs for debugging.

How do I install Testing in Claude Code?

Run `npx skills add nteract/nteract --skill testing -a claude-code`. Or copy the skill folder (.agents/skills/testing in nteract/nteract) into .claude/skills/testing in your project. Claude Code loads it when a task matches its description.

How do I install Testing in Codex?

Run `npx skills add nteract/nteract --skill testing -a codex`. Or copy the skill folder (.agents/skills/testing in nteract/nteract) into .agents/skills/testing in your project. Codex loads it when a task matches its description.

Can I use Testing in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add nteract/nteract --skill testing -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/testing, .gemini/skills/testing, .github/skills/testing and .opencode/skills/testing in your project.

What does Testing need to run?

Going by SKILL.md and its folder, Testing needs the command-line tools its instructions call (cargo, pnpm, pytest, python, pip and deno). Our summary lists: Python 3.

Does Testing access the network?

SKILL.md contains no URLs. Its commands use pip, which can reach the network depending on how they are called. This is read from the text; nothing was executed.

Is Testing safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Testing use?

Testing is published under the BSD-3-Clause licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Testing use?

About 2.2k tokens (SKILL.md is roughly 8.8k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Testing?

Skills that share tags, products or a category with Testing: Python Testing Patterns (jh941213/my-cc-harness, 126 stars), Specx Tests (maksimzayats/specx, 202 stars), Python Testing (macalbert/envilder, 138 stars) and Testing Patterns (softspark/ai-toolkit, 179 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Testing?

nteract (a GitHub organization) maintains it in nteract/nteract, which has 179 GitHub stars. The repository holds 10 skills in this directory. The repository was last updated on October 7, 2026.

Source: nteract/nteract on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.