Agent skill

Testing

by kortix-ai in kortix-ai/suna

A skill your agent uses for every Kortix test task, behavior change, bug fix, refactor, API route change, CLI change, SDK change, browser journey, test failure, coverage question, local benchmark…

Custom licenceAuto-check: notesTesting & QA

Install Testing

skills CLI
$ npx skills add kortix-ai/suna --skill testing -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install kortix-ai/suna testing --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/kortix-ai/suna.git skills-src && mkdir -p .claude/skills && cp -r skills-src/.agents/skills/testing .claude/skills/testing && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
testing
GitHub stars
20k
Token cost
~3.6k tokens
SKILL.md length
1,850 words
Files
4 (incl. references)
Skills in repo
20
Repo updated
First seen
Licence
Custom licence

At a glance

A skill your agent uses for every Kortix test task, behavior change, bug fix, refactor, API route change, CLI change, SDK change, browser journey, test failure, coverage question, local benchmark…

  • Works in 7 steps: Assign or reuse one stable flow ID. → Describe the complete contract in… → Implement the contract through HTTP or a… → …
  • Every Kortix test task
  • SKILL.md covers Select the correct test, Write product flows, Run tests and Prove the result, plus 3 more sections
  • Calls pnpm, bun and terraform

What it does

Testing is an agent skill from kortix-ai/suna. Use for every Kortix test task, behavior change, bug fix, refactor, API route change, CLI change, SDK change, browser journey, test failure, coverage question, local benchmark, or testing infrastructure change. Enforce the single local-first runner, black-box flow contracts, package-local SDK tests, browser-only Playwright tests, and real input/output verification.

Its SKILL.md is about 3.6k tokens, which your agent loads only when the skill is triggered. The skill folder holds 5 other files, including reference files (for example `agents/openai.yaml`, `references/api-latency-baseline.md` and `references/e2e-evaluation.md`).

It sits in Testing & QA, covering Debugging, Failing and flaky tests and Browser testing. It works with Playwright. The repository describes itself as: The open-source AI Operating System.

When your agent uses it

  • Every Kortix test task
  • Behavior change
  • API route change
  • Browser journey

Example prompts

  • “/testing”

Requirements

  • Docker

Workflow steps

7 steps, taken from the first numbered list in SKILL.md.

  1. Assign or reuse one stable flow ID.
  2. Describe the complete contract in tests/spec/end-to-end.md.
  3. Implement the contract through HTTP or a real CLI process.
  4. Write each ctx.step() as one natural-language action and result.
  5. Cover authentication, setup, the action, read-back proof, negative paths, and
  6. List every touched API route in meta.routes.
  7. Regenerate tests/spec/routes.generated.json after route changes.

What it can do on your machine

Read from SKILL.md and the folder at commit 0d0613a. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Shell commands in SKILL.md call:

    • pnpm
    • bun
    • terraform
    • git
    • node
    • supabase

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md. Its commands use pnpm, git and supabase, which can reach the network depending on how they are called.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Testing loads about 3.6k tokens when it runs, and up to ~14k if it reads all its reference files. Until then it costs about 94 tokens; SKILL.md has 1,850 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~94
When it runs · the whole SKILL.md, loaded when a task matches
~3.6k
With references · SKILL.md plus every file in references/, read only if the agent opens them
~14k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check: notes

The automated check noted patterns worth knowing about, such as sudo or a known installer.

  • NoteMentions a .env fileSKILL.md:161
    hook encrypts `.env` files and runs `scripts/check-blocked-terms.sh`.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

Its licence (Custom licence) doesn't allow us to republish the file, so here is its outline and opening line. It has 1,850 words (~3,599 tokens).

“Use one repository-level command: pnpm test.”

— opening of SKILL.md by kortix-ai, Custom licence
name
testing

Read the full SKILL.md on GitHub

Files

SKILL.md and 3 other files (references) in .agents/skills/testing of kortix-ai/suna.

  • SKILL.md
  • agents/openai.yaml
  • references/api-latency-baseline.md
  • references/e2e-evaluation.md

Open the folder on GitHubat commit 0d0613a

Compare with similar skills

Testing next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Testing compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Testing this skillkortix-ai/suna20k—~3.6kAutomated safety check: NotesCustom licence
Debuggingidavidov13/agentic-playwright223—~6.2kAutomated safety check: NotesMIT
E2Esendou-ink/sendou.ink297—~2.1kAutomated safety check: NotesAGPL-3.0
Debug E2E Testbitovi/ai-enablement-prompts121—~677Automated safety check: PassMIT
Activitypub TestingMicrock/ordinary-claude-skills401—~712Automated safety check: PassCustom licence
Playwrightpproenca/dot-skills214—~1.7kAutomated safety check: PassMIT

Similar skills

  • Debugging

    idavidov13/agentic-playwright

    Playwright test debugging conventions for the scaffold — reading failure messages, classifying failure modes (TimeoutError, ZodError, strict-mode violation, locator not found, network errors, schema…

    223 GitHub stars~6.2k tokensUpdated 6 days ago
    Testing & QAAuto-check: notes
  • E2E

    sendou-ink/sendou.ink

    Run, debug, and manage Playwright e2e tests. An agent skill from sendou-ink/sendou.ink.

    297 GitHub stars~2.1k tokensUpdated yesterday
    Testing & QAAuto-check: notes
  • Debug E2E Test

    bitovi/ai-enablement-prompts

    Debug and fix failing Playwright E2E tests. An agent skill from bitovi/ai-enablement-prompts.

    121 GitHub stars~677 tokensUpdated 27 days ago
    Testing & QAAuto-check passed
  • Activitypub Testing

    Microck/ordinary-claude-skills

    Testing patterns for PHPUnit and Playwright E2E tests. An agent skill from Microck/ordinary-claude-skills.

    401 GitHub stars~712 tokensUpdated 1 mo ago
    Testing & QAAuto-check passed
  • Playwright

    pproenca/dot-skills

    Playwright testing best practices for Next.js applications (formerly test-playwright).

    214 GitHub stars~1.7k tokensUpdated 1 mo ago
    Testing & QAAuto-check passed
  • UI Visual Debugging

    NangoHQ/nango

    A skill your agent uses when modifying or visually debugging Nango frontend UI, including packages/webapp, packages/connect-ui, browser interactions, screenshots, and visual regressions.

    13k GitHub stars~1.3k tokensUpdated today
    Testing & QAAuto-check passed

More from kortix-ai/suna

All 20 skills in this repo
  • Ponytail Review

    kortix-ai/suna

    Code review focused exclusively on over-engineering. An agent skill from kortix-ai/suna.

    20k GitHub starsUsed in 4 repos~593 tokens
    Auto-check passed
  • E2E

    kortix-ai/suna

    Agentic end-to-end tests with e2e, the e2e runner. An agent skill from kortix-ai/suna.

    20k GitHub stars~2.3k tokensUpdated today
    Auto-check passed
  • Kortix Brand

    kortix-ai/suna

    Load FIRST for anything that carries the Kortix look or voice: product or mobile UI, copy of any kind, decks, social, images, email, CLI output, anything with the logo, and reviews of these.

    20k GitHub stars~4k tokensUpdated today
    Auto-check passed
  • Contributing

    kortix-ai/suna

    The pull request loop for this repo: branch → commit → verify in your own box (local tests + local stack) → PR into main → demo video recorded with agent-browser on the local stack → gh --attach →…

    20k GitHub stars~3k tokensUpdated today
    Auto-check: warnings
  • Claude Code

    kortix-ai/suna

    Drive Anthropic's Claude Code CLI (claude -p) as a non-interactive coding sub-agent from inside Codex.

    20k GitHub stars~1.2k tokensUpdated today
    Auto-check passed
  • Learnings

    kortix-ai/suna

    The project's episodic memory: a timestamped ledger of rules paid for with real outages and near-misses, one entry per incident.

    20k GitHub stars~1.1k tokensUpdated today
    Auto-check passed

Works with

Questions about Testing

What does Testing do?

A skill your agent uses for every Kortix test task, behavior change, bug fix, refactor, API route change, CLI change, SDK change, browser journey, test failure, coverage question, local benchmark…. Testing is an agent skill from kortix-ai/suna. Use for every Kortix test task, behavior change, bug fix, refactor, API route change, CLI change, SDK change, browser journey, test failure, coverage question, local benchmark, or testing infrastructure change.

When should I use Testing?

Testing fits situations like: every Kortix test task; behavior change; API route change; browser journey.

How do I install Testing in Claude Code?

Run `npx skills add kortix-ai/suna --skill testing -a claude-code`. Or copy the skill folder (.agents/skills/testing in kortix-ai/suna) into .claude/skills/testing in your project. Claude Code loads it when a task matches its description.

How do I install Testing in Codex?

Run `npx skills add kortix-ai/suna --skill testing -a codex`. Or copy the skill folder (.agents/skills/testing in kortix-ai/suna) into .agents/skills/testing in your project. Codex loads it when a task matches its description.

Can I use Testing in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add kortix-ai/suna --skill testing -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/testing, .gemini/skills/testing, .github/skills/testing and .opencode/skills/testing in your project.

What does Testing need to run?

Going by SKILL.md and its folder, Testing needs the command-line tools its instructions call (pnpm, bun, terraform, git, node and supabase). Our summary lists: Docker.

Does Testing access the network?

SKILL.md contains no URLs. Its commands use git, which can reach the network depending on how they are called. This is read from the text; nothing was executed.

Is Testing safe to install?

Our automated static check of SKILL.md found notes only (mentions a .env file), nothing it rates as a warning. It is not a guarantee. Review the folder before installing.

What licence does Testing use?

Testing has a licence file (the repository's licence) that doesn't match a standard licence. Read it on GitHub before reusing the skill.

How many tokens does Testing use?

About 3.6k tokens (SKILL.md is roughly 14k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full. Its references folder adds about 11k tokens, read only when the agent opens those files.

What are the alternatives to Testing?

Skills that share tags, products or a category with Testing: Debugging (idavidov13/agentic-playwright, 223 stars), E2E (sendou-ink/sendou.ink, 297 stars), Debug E2E Test (bitovi/ai-enablement-prompts, 121 stars) and Activitypub Testing (Microck/ordinary-claude-skills, 401 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Testing?

kortix-ai (a GitHub organization) maintains it in kortix-ai/suna, which has 20,256 GitHub stars. The repository holds 20 skills in this directory. The repository was last updated on October 7, 2026.

Source: kortix-ai/suna on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.