Agent skill

Phoenix Pxi Playwright

by Arize-ai in Arize-ai/phoenix

Write, extend, and debug PXI Playwright E2E tests for Phoenix.

Custom licenceAuto-check passedTesting & QA

Install Phoenix Pxi Playwright

skills CLI
$ npx skills add Arize-ai/phoenix --skill phoenix-pxi-playwright -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install Arize-ai/phoenix phoenix-pxi-playwright --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/Arize-ai/phoenix.git skills-src && mkdir -p .claude/skills && cp -r skills-src/.agents/skills/phoenix-pxi-playwright .claude/skills/phoenix-pxi-playwright && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
phoenix-pxi-playwright
GitHub stars
12k
Token cost
~2.6k tokens
SKILL.md length
983 words
Files
1
Skills in repo
39
Repo updated
First seen
Licence
Custom licence

At a glance

Write, extend, and debug PXI Playwright E2E tests for Phoenix.

  • Works in 8 steps: Add the scenario prompt and expected… → Put the scenario prompt, PXI user… → Drive PXI through the real UI with… → …
  • Adding PXI agent frontend specs
  • SKILL.md covers Start Here, Current Harness, Authoring Workflow and Spec Pattern, plus 6 more sections
  • Calls pnpm; needs OPENAI_API_KEY and ANTHROPIC_API_KEY

What it does

Phoenix Pxi Playwright is an agent skill from Arize-ai/phoenix. Write, extend, and debug PXI Playwright E2E tests for Phoenix. Use when adding PXI agent frontend specs, authoring LLM-as-judge rubrics, asserting PXI tool use, persisting PXI test runs as Phoenix experiments, or debugging PXI E2E failures.

Its SKILL.md is about 2.6k tokens, which your agent loads only when the skill is triggered. It is a single SKILL.md file with no bundled scripts.

It sits in Testing & QA, covering LLM observability, Browser testing and LLM evaluation. It works with Playwright. The repository describes itself as: AI Observability & Evaluation.

When your agent uses it

  • Adding PXI agent frontend specs
  • Authoring LLM-as-judge rubrics
  • Asserting PXI tool use
  • Persisting PXI test runs as Phoenix experiments

Example prompts

  • “/phoenix-pxi-playwright”

Requirements

  • A credential in OPENAI_API_KEY
  • A credential in ANTHROPIC_API_KEY

Workflow steps

8 steps, taken from the first numbered list in SKILL.md.

  1. Add the scenario prompt and expected output to PXI_EXPERIMENT_EXAMPLES so the shared PXI E2E dataset gets one example per test scenario.
  2. Put the scenario prompt, PXI user instructions, and judge rubric in the spec file so the test is readable top-to-bottom. Import the…
  3. Drive PXI through the real UI with pxi.open, pxi.acknowledgeConsent, and pxi.askAndWait.
  4. Add deterministic assertions before judge assertions, such as expected text, no agent error, or expected tool use.
  5. Put all post-turn deterministic assertions inside evaluatePxiOutcome, including pxi.expectBackendToolSpanCalled(turn), so failures after…
  6. After pxi.askAndWait returns a turn, persist both passing and failing outcomes. Do not let deterministic assertion failures skip…
  7. Use evaluatePxiOutcome instead of writing per-spec try/catch blocks. It runs judge even when deterministic Playwright assertions fail…
  8. Run the targeted spec with isolated ports before reporting success.

What it can do on your machine

Read from SKILL.md and the folder at commit e471315. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Shell commands in SKILL.md call:

    • pnpm

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md. Its commands use pnpm, which can reach the network depending on how they are called.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names these keys or tokens, usually read from environment variables:

    • OPENAI_API_KEY
    • ANTHROPIC_API_KEY
    • PXI_E2E_EXPERIMENT_BEARER_TOKEN

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Phoenix Pxi Playwright loads about 2.6k tokens when it runs. Until then it costs about 66 tokens; SKILL.md has 983 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~66
When it runs · the whole SKILL.md, loaded when a task matches
~2.6k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

Its licence (Custom licence) doesn't allow us to republish the file, so here is its outline and opening line. It has 983 words (~2,642 tokens).

“Use this skill when authoring or maintaining Playwright specs for PXI, Phoenix's built-in AI assistant. The concrete harness lives in js/app/tests/pxi/; this skill is the authoring guide for using and extending that harness.”

— opening of SKILL.md by Arize-ai, Custom licence
name
phoenix-pxi-playwright
metadata.internal
true

Read the full SKILL.md on GitHub

Files

Just SKILL.md in .agents/skills/phoenix-pxi-playwright of Arize-ai/phoenix.

Open the folder on GitHubat commit e471315

Compare with similar skills

Phoenix Pxi Playwright next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Phoenix Pxi Playwright compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Phoenix Pxi Playwright this skillArize-ai/phoenix12k—~2.6kAutomated safety check: PassCustom licence
Eval Loopjacob-dietle/context-os111—~5.2kAutomated safety check: PassMIT
Web Application Testinganthropics/skills180k51 repos~966Automated safety check: PassApache-2.0
Write and Verify Playwright Testsappsmithorg/appsmith41k—~2.9kAutomated safety check: NotesApache-2.0
playwright-cli Browser Automationgithub/gh-aw5.4k25 repos~2.8kAutomated safety check: PassMIT
Cucumber and Playwright E2E Testslanggenius/dify158k—~682Automated safety check: PassCustom licence

Similar skills

  • Eval Loop

    jacob-dietle/context-os

    This skill should be used when a specific quality problem (UX, data, architecture, feature) needs systematic diagnosis and iterative fixing toward a defined target.

    111 GitHub stars~5.2k tokensUpdated 1 mo ago
    Testing & QAAuto-check passed
  • Web Application Testing

    anthropics/skills

    Official

    Tests local web applications with Python Playwright scripts, checking frontend behavior, capturing screenshots and reading browser console logs.

    180k GitHub starsUsed in 51 repos~966 tokens
    Testing & QAAuto-check passed
  • Writes a Playwright end-to-end test from a prompt, runs it against a live Appsmith deployment and retries with fixes up to three times until it passes.

    41k GitHub stars~2.9k tokensUpdated yesterday
    Testing & QAAuto-check: notes
  • Official

    Drives a real browser from the command line with playwright-cli to open pages, interact, mock requests, save state and work with Playwright tests.

    5.4k GitHub starsUsed in 25 repos~2.8k tokens
    Testing & QAAuto-check passed
  • Guides changes and reviews of the Cucumber and Playwright end-to-end suite under `e2e/`: feature files, step definitions, support code, tags, locators and assertions.

    158k GitHub stars~682 tokensUpdated today
    Testing & QAAuto-check passed
  • E2E Testing

    langflow-ai/langflow

    Write and review Playwright E2E tests for Langflow. An agent skill from langflow-ai/langflow.

    155k GitHub stars~3.3k tokensUpdated yesterday
    Testing & QAAuto-check passed

More from Arize-ai/phoenix

All 39 skills in this repo
  • Harbor Exec

    Arize-ai/phoenix

    A skill your agent uses when working with Harbor's harbor exec CLI workflow: compiling files, directories, or globs into Harbor tasks; running map jobs; configuring artifacts and existence-only…

    12k GitHub stars~909 tokensUpdated yesterday
    Auto-check passed
  • Mintlify

    Arize-ai/phoenix

    Build and maintain documentation sites with Mintlify. An agent skill from Arize-ai/phoenix.

    12k GitHub starsUsed in 8 repos~3.4k tokens
    Auto-check passed
  • Phoenix Frontend

    Arize-ai/phoenix

    Frontend development guidelines for the Phoenix AI observability platform.

    12k GitHub stars~709 tokensUpdated yesterday
    Auto-check passed
  • Phoenix Graphql

    Arize-ai/phoenix

    Write efficient GraphQL queries against the Phoenix API. An agent skill from Arize-ai/phoenix.

    12k GitHub stars~2.2k tokensUpdated yesterday
    Auto-check passed
  • Phoenix Server

    Arize-ai/phoenix

    Backend development guide for the Phoenix AI observability platform (Strawberry GraphQL, SQLAlchemy async, FastAPI).

    12k GitHub stars~1.6k tokensUpdated yesterday
    Auto-check passed
  • Phoenix Storybook

    Arize-ai/phoenix

    Conventions for creating, modifying, and reviewing production-faithful Storybook stories in the Phoenix frontend (js/app/stories, js/app/.storybook).

    12k GitHub stars~1.9k tokensUpdated yesterday
    Auto-check passed

Works with

Questions about Phoenix Pxi Playwright

What does Phoenix Pxi Playwright do?

Write, extend, and debug PXI Playwright E2E tests for Phoenix. Phoenix Pxi Playwright is an agent skill from Arize-ai/phoenix. Write, extend, and debug PXI Playwright E2E tests for Phoenix.

When should I use Phoenix Pxi Playwright?

Phoenix Pxi Playwright fits situations like: adding PXI agent frontend specs; authoring LLM-as-judge rubrics; asserting PXI tool use; persisting PXI test runs as Phoenix experiments.

How do I install Phoenix Pxi Playwright in Claude Code?

Run `npx skills add Arize-ai/phoenix --skill phoenix-pxi-playwright -a claude-code`. Or copy the skill folder (.agents/skills/phoenix-pxi-playwright in Arize-ai/phoenix) into .claude/skills/phoenix-pxi-playwright in your project. Claude Code loads it when a task matches its description.

How do I install Phoenix Pxi Playwright in Codex?

Run `npx skills add Arize-ai/phoenix --skill phoenix-pxi-playwright -a codex`. Or copy the skill folder (.agents/skills/phoenix-pxi-playwright in Arize-ai/phoenix) into .agents/skills/phoenix-pxi-playwright in your project. Codex loads it when a task matches its description.

Can I use Phoenix Pxi Playwright in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add Arize-ai/phoenix --skill phoenix-pxi-playwright -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/phoenix-pxi-playwright, .gemini/skills/phoenix-pxi-playwright, .github/skills/phoenix-pxi-playwright and .opencode/skills/phoenix-pxi-playwright in your project.

What does Phoenix Pxi Playwright need to run?

Going by SKILL.md and its folder, Phoenix Pxi Playwright needs the command-line tools its instructions call (pnpm) and credentials named OPENAI_API_KEY, ANTHROPIC_API_KEY and PXI_E2E_EXPERIMENT_BEARER_TOKEN. Our summary lists: A credential in OPENAI_API_KEY; A credential in ANTHROPIC_API_KEY.

Does Phoenix Pxi Playwright access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is Phoenix Pxi Playwright safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Phoenix Pxi Playwright use?

Phoenix Pxi Playwright has a licence file (the repository's licence) that doesn't match a standard licence. Read it on GitHub before reusing the skill.

How many tokens does Phoenix Pxi Playwright use?

About 2.6k tokens (SKILL.md is roughly 11k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Phoenix Pxi Playwright?

Skills that share tags, products or a category with Phoenix Pxi Playwright: Eval Loop (jacob-dietle/context-os, 111 stars), Web Application Testing (anthropics/skills, 180k stars), Write and Verify Playwright Tests (appsmithorg/appsmith, 41k stars) and playwright-cli Browser Automation (github/gh-aw, 5.4k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Phoenix Pxi Playwright?

Arize-ai (a GitHub organization) maintains it in Arize-ai/phoenix, which has 11,770 GitHub stars. The repository holds 39 skills in this directory. The repository was last updated on October 10, 2026.

Source: Arize-ai/phoenix on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.