Agent skill

QA

by Kiln-AI in Kiln-AI/Kiln

Multi-agent, browser-driven QA pass over a branch or PR — scope it, plan it, fan out one subagent per testing area in its own isolated sandbox, compile a severity-ranked markdown report.

Custom licenceAuto-check passedTesting & QA

Install QA

skills CLI
$ npx skills add Kiln-AI/Kiln --skill qa -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install Kiln-AI/Kiln qa --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/Kiln-AI/Kiln.git skills-src && mkdir -p .claude/skills && cp -r skills-src/.agents/skills/qa .claude/skills/qa && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
qa
GitHub stars
5.2k
Token cost
~3.6k tokens
SKILL.md length
2,128 words
Files
1
Skills in repo
14
Repo updated
First seen
Licence
Custom licence

At a glance

Multi-agent, browser-driven QA pass over a branch or PR — scope it, plan it, fan out one subagent per testing area in its own isolated sandbox, compile a severity-ranked markdown report.

  • Works in 4 steps: Ask scope, then get oriented → Write a plan, get approval before… → Launch one subagent per lane, isolated → …
  • Bug bash request — broader than driving the UI to check one fix (see playwright)
  • SKILL.md covers When to use this vs. something…, Process, Isolating parallel lanes and Briefing each lane's subagent, plus 2 more sections
  • Calls git, bash and npm; needs OPENROUTER_QA_KEY and OPENROUTER_API_KEY

What it does

QA is an agent skill from Kiln-AI/Kiln. Multi-agent, browser-driven QA pass over a branch or PR — scope it, plan it, fan out one subagent per testing area in its own isolated sandbox, compile a severity-ranked markdown report. Use for an "E2E test", "QA pass", "manual test", or "bug bash" request — broader than driving the UI to check one fix (see playwright) or reviewing a diff (see code-review).

Its SKILL.md is about 3.6k tokens, which your agent loads only when the skill is triggered. It is a single SKILL.md file with no bundled scripts.

It sits in Testing & QA, covering End-to-end testing, Browser testing and Subagents. It works with Playwright and Bash. The repository describes itself as: Build, Evaluate, and Optimize AI Systems. Includes evals, RAG, agents, fine-tuning, synthetic data generation, dataset management, MCP, and more.

When your agent uses it

  • Bug bash request — broader than driving the UI to check one fix (see playwright)
  • Reviewing a diff (see code-review)

Example prompts

  • “E2E test”
  • “QA pass”
  • “manual test”
  • “/qa”

Requirements

  • A credential in OPENROUTER_QA_KEY

Workflow steps

4 steps, taken from the step headings in SKILL.md.

  1. Ask scope, then get oriented
  2. Write a plan, get approval before spending anything
  3. Launch one subagent per lane, isolated
  4. Compile the report

What it can do on your machine

Read from SKILL.md and the folder at commit 8ac9075. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Shell commands in SKILL.md call:

    • git
    • bash
    • npm
    • pytest

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md. Its commands use git and npm, which can reach the network depending on how they are called.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names these keys or tokens, usually read from environment variables:

    • OPENROUTER_QA_KEY
    • OPENROUTER_API_KEY

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

QA loads about 3.6k tokens when it runs. Until then it costs about 92 tokens; SKILL.md has 2,128 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~92
When it runs · the whole SKILL.md, loaded when a task matches
~3.6k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

Its licence (Custom licence) doesn't allow us to republish the file, so here is its outline and opening line. It has 2,128 words (~3,574 tokens).

“Driving Kiln's real UI, as a user would, to find what's broken — not code review, not the automated Playwright suite. Read the playwright skill's SKILL.md and references/driving_the_ui.md before doing anything here: this skill is the process wrapper around it…”

— opening of SKILL.md by Kiln-AI, Custom licence
name
qa

Read the full SKILL.md on GitHub

Files

Just SKILL.md in .agents/skills/qa of Kiln-AI/Kiln.

Open the folder on GitHubat commit 8ac9075

Compare with similar skills

QA next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

QA compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
QA this skillKiln-AI/Kiln5.2k—~3.6kAutomated safety check: PassCustom licence
Bangle Followup Operatorbangle-io/bangle-io1.2k—~991Automated safety check: PassAGPL-3.0
Web Application Testinganthropics/skills180k51 repos~966Automated safety check: PassApache-2.0
Write and Verify Playwright Testsappsmithorg/appsmith41k—~2.9kAutomated safety check: NotesApache-2.0
playwright-cli Browser Automationgithub/gh-aw5.4k23 repos~2.8kAutomated safety check: PassMIT
Cucumber and Playwright E2E Testslanggenius/dify158k—~682Automated safety check: PassCustom licence

Similar skills

  • Bangle Followup Operator

    bangle-io/bangle-io

    Operate recurring Bangle.io agent follow-ups after a thread has already started.

    1.2k GitHub stars~991 tokensUpdated 5 days ago
    DevelopmentAuto-check passed
  • Web Application Testing

    anthropics/skills

    Official

    Tests local web applications with Python Playwright scripts, checking frontend behavior, capturing screenshots and reading browser console logs.

    180k GitHub starsUsed in 51 repos~966 tokens
    Testing & QAAuto-check passed
  • Writes a Playwright end-to-end test from a prompt, runs it against a live Appsmith deployment and retries with fixes up to three times until it passes.

    41k GitHub stars~2.9k tokensUpdated today
    Testing & QAAuto-check: notes
  • Official

    Drives a real browser from the command line with playwright-cli to open pages, interact, mock requests, save state and work with Playwright tests.

    5.4k GitHub starsUsed in 23 repos~2.8k tokens
    Testing & QAAuto-check passed
  • Guides changes and reviews of the Cucumber and Playwright end-to-end suite under `e2e/`: feature files, step definitions, support code, tags, locators and assertions.

    158k GitHub stars~682 tokensUpdated today
    Testing & QAAuto-check passed
  • E2E Testing

    langflow-ai/langflow

    Write and review Playwright E2E tests for Langflow. An agent skill from langflow-ai/langflow.

    155k GitHub stars~3.3k tokensUpdated today
    Testing & QAAuto-check passed

More from Kiln-AI/Kiln

All 14 skills in this repo
  • Check Kiln's model list for deprecated or sunset models across all providers.

    5.2k GitHub stars~2.5k tokensUpdated yesterday
    Auto-check: notes
  • Check Kiln's fine-tunable model list for deprecated or unsupported base models.

    5.2k GitHub stars~1.9k tokensUpdated yesterday
    Auto-check: notes
  • Release Digest

    Kiln-AI/Kiln

    Post a "what's changed since the last release" recap to the release Slack channel for final QA.

    5.2k GitHub stars~2.7k tokensUpdated yesterday
    Auto-check passed
  • Kiln Conventions

    Kiln-AI/Kiln

    Kiln's code conventions and per-area gotchas, plus a diff gate that enforces the mechanical ones.

    5.2k GitHub stars~1.7k tokensUpdated yesterday
    Auto-check: notes
  • Playwright

    Kiln-AI/Kiln

    Look at Kiln's UI in a real browser, and run its end-to-end tests.

    5.2k GitHub stars~1.3k tokensUpdated yesterday
    Auto-check passed
  • Open PR

    Kiln-AI/Kiln

    Open a pull request on the Kiln repo for a human to complete.

    5.2k GitHub stars~6.1k tokensUpdated yesterday
    Auto-check passed

Works with

Categories

Questions about QA

What does QA do?

Multi-agent, browser-driven QA pass over a branch or PR — scope it, plan it, fan out one subagent per testing area in its own isolated sandbox, compile a severity-ranked markdown report. QA is an agent skill from Kiln-AI/Kiln. Multi-agent, browser-driven QA pass over a branch or PR — scope it, plan it, fan out one subagent per testing area in its own isolated sandbox, compile a severity-ranked markdown report.

When should I use QA?

QA fits situations like: bug bash request — broader than driving the UI to check one fix (see playwright); reviewing a diff (see code-review).

How do I install QA in Claude Code?

Run `npx skills add Kiln-AI/Kiln --skill qa -a claude-code`. Or copy the skill folder (.agents/skills/qa in Kiln-AI/Kiln) into .claude/skills/qa in your project. Claude Code loads it when a task matches its description.

How do I install QA in Codex?

Run `npx skills add Kiln-AI/Kiln --skill qa -a codex`. Or copy the skill folder (.agents/skills/qa in Kiln-AI/Kiln) into .agents/skills/qa in your project. Codex loads it when a task matches its description.

Can I use QA in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add Kiln-AI/Kiln --skill qa -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/qa, .gemini/skills/qa, .github/skills/qa and .opencode/skills/qa in your project.

What does QA need to run?

Going by SKILL.md and its folder, QA needs the command-line tools its instructions call (git, bash, npm and pytest) and credentials named OPENROUTER_QA_KEY and OPENROUTER_API_KEY. Our summary lists: A credential in OPENROUTER_QA_KEY.

Does QA access the network?

SKILL.md contains no URLs. Its commands use git and npm, which can reach the network depending on how they are called. This is read from the text; nothing was executed.

Is QA safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does QA use?

QA has a licence file (the repository's licence) that doesn't match a standard licence. Read it on GitHub before reusing the skill.

How many tokens does QA use?

About 3.6k tokens (SKILL.md is roughly 14k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to QA?

Skills that share tags, products or a category with QA: Bangle Followup Operator (bangle-io/bangle-io, 1.2k stars), Web Application Testing (anthropics/skills, 180k stars), Write and Verify Playwright Tests (appsmithorg/appsmith, 41k stars) and playwright-cli Browser Automation (github/gh-aw, 5.4k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains QA?

Kiln-AI (a GitHub organization) maintains it in Kiln-AI/Kiln, which has 5,181 GitHub stars. The repository holds 14 skills in this directory. The repository was last updated on October 9, 2026.

Source: Kiln-AI/Kiln on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.