Agent skill

Playground

by Arize-ai in Arize-ai/phoenix

Author, edit, or iterate on prompts in the Phoenix prompt playground, including running experiments over a dataset.

Custom licenceAuto-check passedAI & LLM Engineering

Install Playground

skills CLI
$ npx skills add Arize-ai/phoenix --skill playground -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install Arize-ai/phoenix playground --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/Arize-ai/phoenix.git skills-src && mkdir -p .claude/skills && cp -r skills-src/src/phoenix/server/agents/prompts/skills/playground .claude/skills/playground && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
playground
GitHub stars
12k
Token cost
~3.9k tokens
SKILL.md length
1,738 words
Files
1
Skills in repo
39
Repo updated
First seen
Licence
Custom licence

At a glance

Author, edit, or iterate on prompts in the Phoenix prompt playground, including running experiments over a dataset.

  • Works in 11 steps: Clarify the task the prompt must… → If a playground prompt already exists,… → Draft or revise the prompt so it clearly… → …
  • Tasks that involve LLM observability
  • SKILL.md covers Workflow: Create And Iterate…, Workflow: Iterate Over A… and Workflow: Author, Refine, Or…
  • Instructions only: no scripts, shell commands, URLs or credentials in SKILL.md

What it does

Playground is an agent skill from Arize-ai/phoenix. Author, edit, or iterate on prompts in the Phoenix prompt playground, including running experiments over a dataset. Load before any playground ui. operation call, including single-shot prompt rewrites.

Its SKILL.md is about 3.9k tokens, which your agent loads only when the skill is triggered. It is a single SKILL.md file with no bundled scripts.

It sits in AI & LLM Engineering, covering LLM observability. The repository describes itself as: AI Observability & Evaluation.

When your agent uses it

  • Tasks that involve LLM observability

Example prompts

  • “/playground”

Workflow steps

11 steps, taken from the first numbered list in SKILL.md.

  1. Clarify the task the prompt must perform: input variables, expected output shape, audience,
  2. If a playground prompt already exists, call ui.playground.prompt.read before proposing
  3. Draft or revise the prompt so it clearly states the task, required context, output contract, and
  4. Use ui.playground.prompt.edit for changes to the mounted prompt so the user can review the
  5. Use ui.playground.instance.add when the user wants a fresh comparison instance that starts
  6. Use ui.playground.variables.set when the user provides manual values for prompt template
  7. Use ui.playground.repetitions.set before running when the user is concerned about flakes,
  8. Call ui.playground.run only when the user asks to run, try, test, or compare the current
  9. After the run finishes, call ui.playground.run.readOutput to inspect raw output and get the
  10. Call ui.playground.prompt.save only when the user explicitly asks to save or confirms that the
  11. Inspect the output with the user, identify the next concrete improvement, and repeat the edit or

What it can do on your machine

Read from SKILL.md and the folder at commit 856100b. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    No scripts in the folder and no shell commands in SKILL.md (its code samples are json and javascript).

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Playground loads about 3.9k tokens when it runs. Until then it costs about 54 tokens; SKILL.md has 1,738 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~54
When it runs · the whole SKILL.md, loaded when a task matches
~3.9k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

Its licence (Custom licence) doesn't allow us to republish the file, so here is its outline and opening line. It has 1,738 words (~3,901 tokens).

“The prompt playground is a tool for authoring and optimizing prompts. It supports two different ways of working: fast manual prompt iteration without a dataset, and dataset-backed prompt experimentation with evaluators and experiments. Choose the workflow that matches the user's…”

— opening of SKILL.md by Arize-ai, Custom licence
name
playground
summary
Author, edit, run, compare, and improve prompts in the Phoenix playground.

Read the full SKILL.md on GitHub

Files

Just SKILL.md in src/phoenix/server/agents/prompts/skills/playground of Arize-ai/phoenix.

Open the folder on GitHubat commit 856100b

Compare with similar skills

Playground next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Playground compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Playground this skillArize-ai/phoenix12k—~3.9kAutomated safety check: PassCustom licence
Langfuse Codebase Navigatorlangfuse/langfuse36k—~1.4kAutomated safety check: PassCustom licence
LLM Trace Review Interfaceai-evals-course/evals-skills1.5k—~1.4kAutomated safety check: PassApache-2.0
Langfuse Integration Pagelangfuse/langfuse-docs246—~3.7kAutomated safety check: PassMIT
Langfuselangfuse/skills299—~2.1kAutomated safety check: NotesMIT
Livetablegurujada/live_table211—~1.9kAutomated safety check: PassMIT

Similar skills

  • Navigate Langfuse repositories, code areas, and agent skills.

    36k GitHub stars~1.4k tokensUpdated today
    AI & LLM EngineeringAuto-check passed
  • LLM Trace Review Interface

    ai-evals-course/evals-skills

    Builds a browser-based annotation page for reviewing LLM traces one at a time with pass/fail labels, notes and saved results, tailored to your data.

    1.5k GitHub stars~1.4k tokensUpdated 14 days ago
    AI & LLM EngineeringAuto-check passed
  • Langfuse Integration Page

    langfuse/langfuse-docs

    Create a new Langfuse integration page in the langfuse-docs repo.

    246 GitHub stars~3.7k tokensUpdated today
    AI & LLM EngineeringAuto-check passed
  • Langfuse

    langfuse/skills

    Interact with Langfuse and access its documentation: tracing, monitoring, creating datasets, running experiments, and evaluating AI applications.

    299 GitHub stars~2.1k tokensUpdated 7 days ago
    AI & LLM EngineeringAuto-check: notes
  • Livetable

    gurujada/live_table

    A skill your agent uses when building, modifying, or reviewing Phoenix LiveView tables with LiveTable, including schema-backed tables, context-owned data providers, joined queries, filters…

    211 GitHub stars~1.9k tokensUpdated 3 mo ago
    AI & LLM EngineeringAuto-check passed
  • Langsmith Online Eval Engineering

    langchain-ai/langsmith-skills

    Official

    Design, test, create, and attach LangSmith online evaluators for production traces or conversation threads.

    159 GitHub stars~1.4k tokensUpdated 5 days ago
    AI & LLM EngineeringAuto-check passed

More from Arize-ai/phoenix

All 39 skills in this repo
  • Harbor Exec

    Arize-ai/phoenix

    A skill your agent uses when working with Harbor's harbor exec CLI workflow: compiling files, directories, or globs into Harbor tasks; running map jobs; configuring artifacts and existence-only…

    12k GitHub stars~909 tokensUpdated today
    Auto-check passed
  • Mintlify

    Arize-ai/phoenix

    Build and maintain documentation sites with Mintlify. An agent skill from Arize-ai/phoenix.

    12k GitHub starsUsed in 8 repos~3.4k tokens
    Auto-check passed
  • Phoenix Frontend

    Arize-ai/phoenix

    Frontend development guidelines for the Phoenix AI observability platform.

    12k GitHub stars~709 tokensUpdated today
    Auto-check passed
  • Phoenix Graphql

    Arize-ai/phoenix

    Write efficient GraphQL queries against the Phoenix API. An agent skill from Arize-ai/phoenix.

    12k GitHub stars~2.2k tokensUpdated today
    Auto-check passed
  • Phoenix Server

    Arize-ai/phoenix

    Backend development guide for the Phoenix AI observability platform (Strawberry GraphQL, SQLAlchemy async, FastAPI).

    12k GitHub stars~1.6k tokensUpdated today
    Auto-check passed
  • Phoenix Storybook

    Arize-ai/phoenix

    Conventions for creating, modifying, and reviewing production-faithful Storybook stories in the Phoenix frontend (js/app/stories, js/app/.storybook).

    12k GitHub stars~1.9k tokensUpdated today
    Auto-check passed

Questions about Playground

What does Playground do?

Author, edit, or iterate on prompts in the Phoenix prompt playground, including running experiments over a dataset. Playground is an agent skill from Arize-ai/phoenix. Author, edit, or iterate on prompts in the Phoenix prompt playground, including running experiments over a dataset.

When should I use Playground?

Playground fits situations like: tasks that involve LLM observability.

How do I install Playground in Claude Code?

Run `npx skills add Arize-ai/phoenix --skill playground -a claude-code`. Or copy the skill folder (src/phoenix/server/agents/prompts/skills/playground in Arize-ai/phoenix) into .claude/skills/playground in your project. Claude Code loads it when a task matches its description.

How do I install Playground in Codex?

Run `npx skills add Arize-ai/phoenix --skill playground -a codex`. Or copy the skill folder (src/phoenix/server/agents/prompts/skills/playground in Arize-ai/phoenix) into .agents/skills/playground in your project. Codex loads it when a task matches its description.

Can I use Playground in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add Arize-ai/phoenix --skill playground -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/playground, .gemini/skills/playground, .github/skills/playground and .opencode/skills/playground in your project.

What does Playground need to run?

SKILL.md names no scripts, command-line tools or credentials: Playground is instructions for the agent only.

Does Playground access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is Playground safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Playground use?

Playground has a licence file (the repository's licence) that doesn't match a standard licence. Read it on GitHub before reusing the skill.

How many tokens does Playground use?

About 3.9k tokens (SKILL.md is roughly 16k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Playground?

Skills that share tags, products or a category with Playground: Langfuse Codebase Navigator (langfuse/langfuse, 36k stars), LLM Trace Review Interface (ai-evals-course/evals-skills, 1.5k stars), Langfuse Integration Page (langfuse/langfuse-docs, 246 stars) and Langfuse (langfuse/skills, 299 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Playground?

Arize-ai (a GitHub organization) maintains it in Arize-ai/phoenix, which has 11,744 GitHub stars. The repository holds 39 skills in this directory. The repository was last updated on October 8, 2026.

Source: Arize-ai/phoenix on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.