Agent skill

Testing E2E

by yonatangross in yonatangross/orchestkit

End-to-end testing patterns with Playwright — page objects, AI agent testing, visual regression, accessibility testing with axe-core, and CI integration.

MITAuto-check passedTesting & QA

Install Testing E2E

skills CLI
$ npx skills add yonatangross/orchestkit --skill testing-e2e -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install yonatangross/orchestkit testing-e2e --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/yonatangross/orchestkit.git skills-src && mkdir -p .claude/skills && cp -r skills-src/src/skills/testing-e2e .claude/skills/testing-e2e && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
testing-e2e
GitHub stars
288
Token cost
~2.9k tokens
SKILL.md length
977 words
Files
17 (incl. scripts, references)
Skills in repo
107
Repo updated
First seen
Licence
MIT

At a glance

End-to-end testing patterns with Playwright — page objects, AI agent testing, visual regression, accessibility testing with axe-core, and CI integration.

  • Writing E2E tests
  • SKILL.md covers Quick Reference, Upstream coverage (do not…, emulate Backends and Playwright Quick Start, plus 13 more sections
  • Calls npx
  • Setting up Playwright

What it does

Testing E2E is an agent skill from yonatangross/orchestkit. End-to-end testing patterns with Playwright — page objects, AI agent testing, visual regression, accessibility testing with axe-core, and CI integration. Use when writing E2E tests, setting up Playwright, implementing visual regression, or testing accessibility.

Its SKILL.md is about 2.9k tokens, which your agent loads only when the skill is triggered. The skill folder holds 21 other files, including scripts and reference files (for example `checklists/a11y-testing-checklist.md`, `checklists/e2e-checklist.md` and `checklists/e2e-testing-checklist.md`). Compatibility notes: Claude Code 2.1.277+.

It sits in Testing & QA, covering End-to-end testing, Accessibility and Browser testing. It works with Playwright. The repository describes itself as: The Complete AI Development Toolkit for Claude Code. 106 skills, 36 agents, 171 hooks. Install ork for stable (v9.x), or ork-alpha for the v10 line, which ships daily. The licence is MIT.

When your agent uses it

  • Writing E2E tests
  • Setting up Playwright
  • Implementing visual regression
  • Testing accessibility

Example prompts

  • “/testing-e2e”

Requirements

  • Node.js
  • Compatibility (from SKILL.md): Claude Code 2.1.277+.
  • Pre-approved tools (allowed-tools): Read, Glob, Grep, WebFetch, WebSearch

What it can do on your machine

Read from SKILL.md and the folder at commit 1f8d8f3. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves these tools, so the agent can use them without asking each time:

    • Read
    • Glob
    • Grep
    • WebFetch
    • WebSearch

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Ships 1 file in scripts/, which the agent can run.

    Shell commands in SKILL.md call:

    • npx

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    Links to these hosts (documentation or services it may open):

    • playwright.dev
    • github.com
    • w3.org

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

  • Compatibility

    Claude Code 2.1.277+.

    From compatibility in the SKILL.md frontmatter.

Context cost

Testing E2E loads about 2.9k tokens when it runs, and up to ~4.6k if it reads all its reference files. Until then it costs about 69 tokens; SKILL.md has 977 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~69
When it runs · the whole SKILL.md, loaded when a task matches
~2.9k
With references · SKILL.md plus every file in references/, read only if the agent opens them
~4.6k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); the scripts in this folder are not scanned.

SKILL.md

The full file from yonatangross/orchestkit at commit 1f8d8f3, republished under its MIT licence (© yonatangross). 977 words, ~2,945 tokens.

Download SKILL.mdSave it as .claude/skills/testing-e2e/SKILL.md (or your agent's skills folder). This skill also uses 16 other files; get the full folder from GitHub.
name
testing-e2e
description
End-to-end testing patterns with Playwright — page objects, AI agent testing, visual regression, accessibility testing with axe-core, and CI integration. Use when writing E2E tests, setting up Playwright, implementing visual regression, or testing accessibility.
allowed-tools
Read, Glob, Grep, WebFetch, WebSearch
compatibility
Claude Code 2.1.277+.
license
MIT
context
fork
agent
test-generator
user-invocable
false
disable-model-invocation
false
metadata.category
document-asset-creation
metadata.version
2.1.0
metadata.author
OrchestKit
metadata.complexity
medium
metadata.tags
testing, e2e, playwright, accessibility, visual-regression, page-objects

E2E Testing Patterns

End-to-end testing with Playwright 1.59+, visual regression, accessibility, and AI agent workflows.

Quick Reference

CategoryRulesImpactWhen to Use
emulate Backendsrules/emulate-e2e.mdHIGHFIRST CHOICE — deterministic API backends for E2E
Playwright Corerules/e2e-playwright.mdHIGHSemantic locators, auto-wait, flaky detection
Page Objectsrules/e2e-page-objects.mdHIGHEncapsulate page interactions, visual regression
AI Agentsrules/e2e-ai-agents.mdHIGHPlanner/Generator/Healer, init-agents
A11y Playwrightrules/a11y-playwright.mdMEDIUMFull-page axe-core scanning with WCAG 2.2 AA
A11y CI/CDrules/a11y-testing.mdMEDIUMCI gates, jest-axe unit tests, PR blocking
End-to-End Typesrules/validation-end-to-end.mdHIGHtRPC, Prisma, Pydantic type safety

Total: 7 rules, 2 references, 3 checklists, 1 example, 1 script

Upstream coverage (do not restate)

Playwright, axe-core and jest-axe document themselves. This skill carries only the OrchestKit delta (references/ork-delta.md) plus the house subsets in rules/. Fetch the vendor page for anything below instead of expecting it here.

TopicSource
Screenshot comparison workflow, baseline files, snapshotPathTemplatehttps://playwright.dev/docs/test-snapshots
toHaveScreenshot options: mask, maxDiffPixelRatio, stylePath, animationshttps://playwright.dev/docs/api/class-pageassertions
Auth reuse: storageState, setup projects, IndexedDBhttps://playwright.dev/docs/auth
Network interception with page.route (the house default is still emulate first, see rules/emulate-e2e.md)https://playwright.dev/docs/mock
Removed and changed APIs per release (SKILL.md keeps only the short denylist below)https://playwright.dev/docs/release-notes
Full locator API surface (the house priority ladder stays in rules/e2e-playwright.md)https://playwright.dev/docs/locators
Playwright runner setup on CI (the house a11y gate workflow stays in rules/a11y-testing.md)https://playwright.dev/docs/ci
init-agents CLI flags and generated files (the Planner/Generator/Healer workflow stays in rules/e2e-ai-agents.md)https://playwright.dev/docs/test-agents
jest-axe matcher and configureAxe API (the house component-state subset stays in rules/a11y-testing.md)https://github.com/NickColley/jest-axe
Lighthouse CI configuration and score assertionshttps://github.com/GoogleChrome/lighthouse-ci

WCAG 2.2 success criteria and the manual keyboard / screen-reader / contrast / zoom passes are NOT routed away: checklists/a11y-testing-checklist.md still carries them in full, with https://www.w3.org/WAI/WCAG22/quickref/ as the normative reference.

emulate Backends

For E2E tests that interact with external APIs (GitHub, Vercel, Google), use emulate as the backend instead of hitting real APIs. This eliminates flakiness from rate limits, network issues, and non-deterministic data.

ApproachResult
emulate backends (FIRST CHOICE)Deterministic, fast, CI-friendly
Real APIsFlaky, rate-limited, slow
MSW/Nock interceptsNo state machines, manual response management

Key features: seed config for reproducible data, per-worker port isolation for parallel Playwright, full state machine transitions.

See rules/emulate-e2e.md for patterns, CI configuration, and per-worker isolation fixtures.


Playwright Quick Start

typescript
import { test, expect } from '@playwright/test';

test('user can complete checkout', async ({ page }) => {
  await page.goto('/products');
  await page.getByRole('button', { name: 'Add to cart' }).click();
  await page.getByRole('link', { name: 'Checkout' }).click();
  await page.getByLabel('Email').fill('test@example.com');
  await page.getByRole('button', { name: 'Submit' }).click();
  await expect(page.getByRole('heading', { name: 'Order confirmed' })).toBeVisible();
});

Locator Priority: getByRole() > getByLabel() > getByPlaceholder() > getByTestId()

Playwright Core

Semantic locator patterns and best practices for resilient tests.

RuleFileKey Pattern
Playwright E2Erules/e2e-playwright.mdSemantic locators, auto-wait, new 1.58+ features

Anti-patterns (FORBIDDEN):

  • Hardcoded waits: await page.waitForTimeout(2000)
  • CSS selectors for interactions: await page.click('.submit-btn')
  • XPath locators

Removed in 1.58/1.59 — do NOT use:

  • _react=ComponentName[prop=value] and _vue=... component selector engines — removed in 1.58
  • :light selector suffix — removed
  • launch({ devtools: true }) option — removed; use args: ['--auto-open-devtools-for-tabs']

Page Objects

Encapsulate page interactions into reusable classes.

RuleFileKey Pattern
Page Object Modelrules/e2e-page-objects.mdLocators in constructor, action methods, assertion methods
typescript
const checkout = new CheckoutPage(page);
await checkout.fillEmail('test@example.com');
await checkout.submit();
await checkout.expectConfirmation();

AI Agents

Playwright 1.59+ AI agent framework for test planning, generation, and self-healing. Includes a token-efficient CLI mode designed for coding agents — minimal output, structured responses, reduced context overhead.

RuleFileKey Pattern
AI Agentsrules/e2e-ai-agents.mdPlanner, Generator, Healer workflow
bash
npx playwright init-agents --loop=claude    # For Claude Code

Token-efficient CLI mode (1.58+): Playwright ships a SKILL-focused CLI mode that produces compact, agent-friendly output — use this when running Playwright from AI agents to minimize token consumption.

Workflow: Planner (explores app, creates specs) -> Generator (reads spec, tests live app) -> Healer (fixes failures, updates selectors).

New in Playwright 1.59 (Apr 2026) — relevant for AI agents:

  • page.screencast({ start, stop, showActions }) — unified video + real-time JPEG frame streaming. Lets a Healer agent read frames mid-run for visual assertion without writing video files.
  • browser.bind() / npx playwright-cli attach — attach to a running browser from an MCP client mid-test; useful for Healer to inspect a hung or failing CI run.
  • locator.normalize() — rewrites a brittle locator to best-practice equivalents. Pair with Healer to auto-upgrade getByTestId → getByRole where possible.
Show full SKILL.md (362 more words)Show less

Accessibility (Playwright)

Full-page accessibility validation with axe-core in E2E tests.

RuleFileKey Pattern
Playwright + axerules/a11y-playwright.mdWCAG 2.2 AA, interactive state testing
typescript
import AxeBuilder from '@axe-core/playwright';

test('page meets WCAG 2.2 AA', async ({ page }) => {
  await page.goto('/');
  const results = await new AxeBuilder({ page })
    .withTags(['wcag2a', 'wcag2aa', 'wcag22aa'])
    .analyze();
  expect(results.violations).toEqual([]);
});

Accessibility (CI/CD)

CI pipeline integration and jest-axe unit-level component testing.

RuleFileKey Pattern
CI Gates + jest-axerules/a11y-testing.mdPR blocking, component state testing

End-to-End Types

Type safety across API layers to eliminate runtime type errors.

RuleFileKey Pattern
Type Safetyrules/validation-end-to-end.mdtRPC, Zod, Pydantic, schema rejection tests

Visual Regression

Native Playwright screenshot comparison without external services.

typescript
await expect(page).toHaveScreenshot('checkout-page.png', {
  maxDiffPixels: 100,
  mask: [page.locator('.dynamic-content')],
});

House rules that the vendor docs do not state (CI-only baselines, single snapshot project, mask over threshold): references/ork-delta.md. Option reference and the baseline workflow itself: see the Upstream coverage table above.

Key Decisions

DecisionRecommendation
E2E frameworkPlaywright 1.59+ with semantic locators
Locator strategygetByRole > getByLabel > getByTestId
BrowserChromium (Chrome for Testing in 1.59+)
Page patternPage Object Model for complex pages
Visual regressionPlaywright native toHaveScreenshot()
A11y testingaxe-core (E2E) + jest-axe (unit)
CI retries2-3 in CI, 0 locally
Flaky detectionfailOnFlakyTests: true in CI
AI agentsPlanner/Generator/Healer via init-agents
Type safetytRPC for end-to-end, Zod for runtime validation

References

ResourceDescription
references/ork-delta.mdHouse rules the vendor docs do not state: jest-axe over vitest-axe, CLI-only agent init, CI-only baselines, single snapshot project, mask over threshold
references/playwright-setup.mdInstallation, MCP server, seed tests, agent initialization

Checklists

ChecklistDescription
checklists/e2e-checklist.mdLocator strategy, page objects, CI/CD, visual regression
checklists/e2e-testing-checklist.mdComprehensive: planning, implementation, SSE, responsive, maintenance
checklists/a11y-testing-checklist.mdAutomated + manual: keyboard, screen reader, color contrast, WCAG

Examples

ExampleDescription
examples/orchestkit-e2e-tests.mdOrchestKit analysis flow: page objects, SSE progress, error handling

Generic Playwright samples (user flows, auth fixtures, API mocking, multi-tab, file upload, axe scans) now come from the vendor pages in the Upstream coverage table.

Scripts

ScriptDescription
scripts/create-page-object.mdGenerate Playwright page object with auto-detected patterns
  • testing-unit - Unit testing patterns with mocking, fixtures, and data factories
  • testing-integration - API boundary and contract testing
  • cover - Generates the E2E tier when the suite does not exist yet
  • verify - Grades an existing suite and returns a merge verdict
  • expect - Diff-aware browser verification via agent-browser
  • emulate-seed - Seed configuration authoring for emulate providers
  • portless (upstream) - Stable HTTPS baseURL for local E2E tests (https://myapp.localhost instead of port guessing; HTTPS-on-443 default since portless 0.10)

© yonatangross, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

SKILL.md and 16 other files (scripts, references) in src/skills/testing-e2e of yonatangross/orchestkit.

  • SKILL.md
  • checklists/a11y-testing-checklist.md
  • checklists/e2e-checklist.md
  • checklists/e2e-testing-checklist.md
  • examples/orchestkit-e2e-tests.md
  • references/ork-delta.md
  • references/playwright-setup.md
  • rules/_sections.md
  • rules/a11y-playwright.md
  • rules/a11y-testing.md
  • rules/e2e-ai-agents.md
  • rules/e2e-page-objects.md
  • rules/e2e-playwright.md
  • rules/emulate-e2e.md
  • rules/validation-end-to-end.md
  • scripts/create-page-object.md
  • … and 1 more

Open the folder on GitHubat commit 1f8d8f3

Compare with similar skills

Testing E2E next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Testing E2E compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Testing E2E this skillyonatangross/orchestkit288—~2.9kAutomated safety check: PassMIT
Playwright Testingchongdashu/vibejam-starter-pack149—~2.2kAutomated safety check: PassNone
Web Testing with Playwright and Vitestwithkynam/vibecode-pro-max-kit1.1k—~892Automated safety check: PassApache-2.0
Playwright Automationpetrkindlmann/qa-skills163—~5.5kAutomated safety check: PassMIT
Playwright UI TestingHack23/cia239—~2.5kAutomated safety check: PassApache-2.0
Playwright Testingchongdashu/vibejam-starter-pack149—~2.1kAutomated safety check: PassNone

Similar skills

  • Playwright Testing

    chongdashu/vibejam-starter-pack

    Plan, implement, and debug frontend tests: unit/integration/E2E/visual/a11y.

    149 GitHub stars~2.2k tokensUpdated 5 mo ago
    Testing & QAAuto-check passed
  • Web Testing with Playwright and Vitest

    withkynam/vibecode-pro-max-kit

    Covers web testing from unit to E2E, load, visual, accessibility and security checks, with Playwright, Vitest and k6 guides plus a Playwright setup script.

    1.1k GitHub stars~892 tokensUpdated 3 mo ago
    Testing & QAAuto-check passed
  • Playwright Automation

    petrkindlmann/qa-skills

    Write production-grade Playwright tests in TypeScript: Page Object Model, fixtures, auto-waiting, user-facing locators, parallel execution, CI integration, sharding, and 2025-2026 feature awareness.

    163 GitHub stars~5.5k tokensUpdated 3 mo ago
    Testing & QAAuto-check passed
  • Playwright browser automation, visual regression testing, accessibility testing, and E2E workflow validation for CIA platform

    239 GitHub stars~2.5k tokensUpdated today
    Testing & QAAuto-check passed
  • Playwright Testing

    chongdashu/vibejam-starter-pack

    Plan, implement, and debug frontend tests: unit/integration/E2E/visual/a11y.

    149 GitHub stars~2.1k tokensUpdated 5 mo ago
    Testing & QAAuto-check passed
  • UI Visual Debugging

    NangoHQ/nango

    A skill your agent uses when modifying or visually debugging Nango frontend UI, including packages/webapp, packages/connect-ui, browser interactions, screenshots, and visual regressions.

    13k GitHub stars~1.3k tokensUpdated today
    Testing & QAAuto-check passed

More from yonatangross/orchestkit

All 107 skills in this repo
  • API Design

    yonatangross/orchestkit

    API contract design for REST and GraphQL, covering resource shape, URL and header versioning with deprecation windows, RFC 9457 Problem Details error handling, and OpenAPI specs.

    288 GitHub stars~2.9k tokensUpdated yesterday
    Auto-check passed
  • Architecture Decision Record

    yonatangross/orchestkit

    ADR templates in the Nygard format with context, decision, consequences, and alternatives.

    288 GitHub stars~2k tokensUpdated yesterday
    Auto-check passed
  • Audit Full

    yonatangross/orchestkit

    Single-pass codebase analysis leveraging a 1M-token context window for comprehensive security scanning, architecture review, and dependency auditing.

    288 GitHub stars~3.5k tokensUpdated yesterday
    Auto-check: notes
  • Code Review Playbook

    yonatangross/orchestkit

    Structured review processes, conventional comments, language-specific checklists, and feedback templates.

    288 GitHub stars~2.2k tokensUpdated yesterday
    Auto-check passed
  • Create PR

    yonatangross/orchestkit

    Creates GitHub pull requests with pre-flight validation, conventional title formatting, and structured summary generation.

    288 GitHub stars~4.5k tokensUpdated yesterday
    Auto-check: notes
  • Explore

    yonatangross/orchestkit

    Multi-angle codebase exploration spawning 3-5 parallel agents for code structure, data flow, architecture patterns, and health assessment.

    288 GitHub stars~3.9k tokensUpdated yesterday
    Auto-check: notes

Works with

Categories

Questions about Testing E2E

What does Testing E2E do?

End-to-end testing patterns with Playwright — page objects, AI agent testing, visual regression, accessibility testing with axe-core, and CI integration. Testing E2E is an agent skill from yonatangross/orchestkit. End-to-end testing patterns with Playwright — page objects, AI agent testing, visual regression, accessibility testing with axe-core, and CI integration.

When should I use Testing E2E?

Testing E2E fits situations like: writing E2E tests; setting up Playwright; implementing visual regression; testing accessibility.

How do I install Testing E2E in Claude Code?

Run `npx skills add yonatangross/orchestkit --skill testing-e2e -a claude-code`. Or copy the skill folder (src/skills/testing-e2e in yonatangross/orchestkit) into .claude/skills/testing-e2e in your project. Claude Code loads it when a task matches its description.

How do I install Testing E2E in Codex?

Run `npx skills add yonatangross/orchestkit --skill testing-e2e -a codex`. Or copy the skill folder (src/skills/testing-e2e in yonatangross/orchestkit) into .agents/skills/testing-e2e in your project. Codex loads it when a task matches its description.

Can I use Testing E2E in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add yonatangross/orchestkit --skill testing-e2e -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/testing-e2e, .gemini/skills/testing-e2e, .github/skills/testing-e2e and .opencode/skills/testing-e2e in your project.

What does Testing E2E need to run?

Going by SKILL.md and its folder, Testing E2E needs the command-line tools its instructions call (npx). Our summary lists: Node.js. Its frontmatter pre-approves these tools: Read, Glob, Grep, WebFetch, WebSearch. Compatibility (from SKILL.md): Claude Code 2.1.277+..

Does Testing E2E access the network?

SKILL.md names 3 domains. As links in the text: playwright.dev, github.com and w3.org. This is read from the text; nothing was executed.

Is Testing E2E safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. The check reads SKILL.md only: the scripts in the folder are not scanned, so read them before running anything.

What licence does Testing E2E use?

Testing E2E is published under the MIT licence (declared in SKILL.md). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Testing E2E use?

About 2.9k tokens (SKILL.md is roughly 12k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full. Its references folder adds about 1.7k tokens, read only when the agent opens those files.

What are the alternatives to Testing E2E?

Skills that share tags, products or a category with Testing E2E: Playwright Testing (chongdashu/vibejam-starter-pack, 149 stars), Web Testing with Playwright and Vitest (withkynam/vibecode-pro-max-kit, 1.1k stars), Playwright Automation (petrkindlmann/qa-skills, 163 stars) and Playwright UI Testing (Hack23/cia, 239 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Testing E2E?

yonatangross (a GitHub user) maintains it in yonatangross/orchestkit, which has 288 GitHub stars. The repository holds 107 skills in this directory. The repository was last updated on October 6, 2026.

Source: yonatangross/orchestkit on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.