Agent skill

Control UI E2E

by openclaw in openclaw/openclaw

A skill your agent uses when designing, testing, fixing, or extending the OpenClaw Control UI GUI, including UI stress-test galleries with feedback inputs, Vitest + Playwright end-to-end checks…

MITAuto-check passedTesting & QA

Install Control UI E2E

skills CLI
$ npx skills add openclaw/openclaw --skill control-ui-e2e -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install openclaw/openclaw control-ui-e2e --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/openclaw/openclaw.git skills-src && mkdir -p .claude/skills && cp -r skills-src/.agents/skills/control-ui-e2e .claude/skills/control-ui-e2e && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
control-ui-e2e
GitHub stars
392k
Token cost
~2.9k tokens
SKILL.md length
1,384 words
Files
2
Skills in repo
93
Repo updated
First seen
Licence
MIT

At a glance

A skill your agent uses when designing, testing, fixing, or extending the OpenClaw Control UI GUI, including UI stress-test galleries with feedback inputs, Vitest + Playwright end-to-end checks…

  • Works in 5 steps: Derive examples from the affected… → Build one browser-openable HTML overview… → Give every example a labeled feedback… → …
  • Extending the OpenClaw Control UI GUI
  • SKILL.md covers UI Stress Test, Test Shape, Commands and Visual Proof Default, plus 3 more sections
  • Calls pnpm, node and playwright

What it does

Control UI E2E is an agent skill from openclaw/openclaw. Use when designing, testing, fixing, or extending the OpenClaw Control UI GUI, including UI stress-test galleries with feedback inputs, Vitest + Playwright end-to-end checks, mocked Gateway flows, screenshots/videos, or agent-verifiable browser proof.

Its SKILL.md is about 2.9k tokens, which your agent loads only when the skill is triggered. The skill folder holds 2 other files (for example `agents/openai.yaml`).

It sits in Testing & QA, covering End-to-end testing, Load testing and Unit testing. It works with Playwright and Vitest. The repository describes itself as: The AI that really does things. Any OS. Any Platform. The lobster way. 🦞. The licence is MIT.

When your agent uses it

  • Extending the OpenClaw Control UI GUI
  • Including UI stress-test galleries with feedback inputs
  • Vitest + Playwright end-to-end checks
  • Mocked Gateway flows

Example prompts

  • “/control-ui-e2e”

Workflow steps

5 steps, taken from the first numbered list in SKILL.md.

  1. Derive examples from the affected components and their data contracts. Cover
  2. Build one browser-openable HTML overview with stable example IDs, short state
  3. Give every example a labeled feedback input. Persist feedback locally across
  4. Open the gallery in the available preview browser and share its URL or file
  5. Before requesting feedback, inspect the rendered examples at relevant viewport

What it can do on your machine

Read from SKILL.md and the folder at commit 1eb5970. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Shell commands in SKILL.md call:

    • pnpm
    • node
    • playwright

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    Links to these hosts (documentation or services it may open):

    • docs.openclaw.ai

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Control UI E2E loads about 2.9k tokens when it runs. Until then it costs about 67 tokens; SKILL.md has 1,384 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~67
When it runs · the whole SKILL.md, loaded when a task matches
~2.9k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from openclaw/openclaw at commit 1eb5970, republished under its MIT licence (© openclaw). 1,384 words, ~2,859 tokens.

Download SKILL.mdSave it as .claude/skills/control-ui-e2e/SKILL.md (or your agent's skills folder). This skill also uses 1 other file; get the full folder from GitHub.
name
control-ui-e2e
description
Use when designing, testing, fixing, or extending the OpenClaw Control UI GUI, including UI stress-test galleries with feedback inputs, Vitest + Playwright end-to-end checks, mocked Gateway flows, screenshots/videos, or agent-verifiable browser proof.

Control UI E2E

Use this for Control UI design feedback and real browser flows with deterministic Gateway data.

UI Stress Test

For substantial UI changes, build a local HTML stress-test gallery early so the user can compare meaningful states and give feedback against concrete examples. Use it when changing layouts, interactions, or components with multiple states; small copy or icon edits can skip it when a gallery adds no useful comparison.

  1. Derive examples from the affected components and their data contracts. Cover the relevant normal, loading, empty, error, unavailable, permission, selected, and expanded states, plus long text or dense content where they stress the layout. Show mutually exclusive states separately; label proposed states that the current implementation does not support.
  2. Build one browser-openable HTML overview with stable example IDs, short state labels, and enough context to understand each example. Prefer real components and deterministic mock fixtures. Label static or approximate renderings and link to the running UI for interactions they cannot reproduce. Keep generated galleries in task-owned artifact storage rather than committing them by default.
  3. Give every example a labeled feedback input. Persist feedback locally across refreshes using a gallery-specific storage key and stable example IDs. Include a Copy feedback action that exports the example IDs, state labels, and comments as Markdown or plain text for the user to return to the conversation.
  4. Open the gallery in the available preview browser and share its URL or file path. Keep the same gallery and example IDs during iteration, preserving existing comments as examples change. Apply the user's feedback to both the gallery and the implementation so they remain comparable.
  5. Before requesting feedback, inspect the rendered examples at relevant viewport sizes and verify that feedback survives a refresh and exports with the correct example labels. Report any unsupported states or preview limitations.

The gallery supports design review; it does not replace focused behavior tests or inspected before/after proof from the running UI. Preserve final proof using the fresh capture directories described below, separately from the evolving gallery.

Test Shape

  • Use ui/src/**/*.e2e.test.ts for full GUI flows.
  • Use ui/src/test-helpers/control-ui-e2e.ts to start the Vite Control UI and install a mocked Gateway WebSocket.
  • Keep scenarios deterministic. Do not use live provider keys, real channel credentials, or a real Gateway unless the user explicitly asks for live proof.
  • Prefer existing .browser.test.ts or unit tests for narrow rendering logic; use this E2E lane when the proof should cover routing, app boot, Gateway handshake, requests, and visible UI behavior together.

Commands

  • Target one E2E test in a Codex worktree:
bash
node scripts/run-vitest.mjs run --config test/vitest/vitest.ui-e2e.config.ts --configLoader runner ui/src/e2e/chat-flow.messaging.e2e.test.ts
  • Run the whole local lane in a normal checkout:
bash
pnpm test:ui:e2e

Use an existing ready dependency installation or a prepared normal checkout; do not reconcile a shared install while other jobs use it. Follow $openclaw-testing: trusted development proof may run locally, and remote proof needs a browser/platform, clean-environment, or source-isolation reason.

Visual Proof Default

For appearance changes, capture inspected before/after visual evidence. For other behavior, use the clearest appropriate boundary proof; a video and a screenshot set are not mandatory when assertions already demonstrate the change.

  • Keep the Vitest E2E assertions deterministic; do not commit generated screenshots or videos.
  • The shared suite disables Chromium partial rasterization and GPU rasterization to avoid observed paint-history-dependent edge pixels. Exact repeat comparisons are still required before claiming reproducibility.
  • For transient states such as Saved, install page.clock before the fixture, pause it with pauseVirtualClock, and advance only the fixture work needed to enter that state. setFixedTime alone does not pause timers. Capture readiness uses native layout delivery without advancing the fixture clock.
  • For stills, use takeControlUiScreenshotFrame from ui/src/test-helpers/control-ui-e2e-screenshot.ts with explicit semantic content and animations: "disabled". Pass viewport changes and the intended scroll target through its viewport/scrollTo options; it verifies retained centering, viewport/scroll-clip intersection, and settled layout, fonts, and visible images.
  • Pass all related element locators in elements, then save the returned page PNG and crops from that single frame. Crops enclose fractional bounds in measured PNG pixels; dimensions and unchanged bounds are asserted. Do not combine separate page and locator screenshots as same-frame evidence.
  • The default current-frame mode and existing viewport/element helpers preserve recording and sampled animation behavior. Static preparation stays active through bounds measurement and the unclipped capture; full-page proofs must fit the viewport-frame dimensions.
  • After or alongside the focused E2E test, run the mocked Control UI app when available, for example pnpm dev:ui:mock -- --port <port>.
  • Drive Chromium with Playwright against the local mock URL. Capture the states needed to demonstrate the change, using screenshots or a short video.
  • Use browser.newContext({ recordVideo: { dir, size }, viewport }), page.screenshot({ path }), and close the context before reporting the video path.
  • The session-host command-state proof uses viewport-only captures, verified with Playwright 1.62.1 and Chrome 151.0.7922.34 (Linux real Gateway; macOS arm64 synthetic reproduction). Other recording owners have not been migrated or certified by this fix; verify their required screenshot content and finalized video separately. See the verified capture path and upstream limitation.
  • Allocate retained proof with createControlUiE2eArtifactDir(scope, parentDir?) from ui/src/test-helpers/control-ui-e2e-artifacts.ts. Each call atomically creates a fresh directory and logs its actual path. An explicit parent wins, then the trimmed existing OPENCLAW_UI_E2E_ARTIFACT_DIR, then the repository's .artifacts/control-ui-e2e parent. Existing custom output controls select parents; do not add or rewrite env vars to enable capture.
  • Allocate during the test/scenario or beforeEach, once per attempt; standalone scripts allocate once per invocation. Pass the owner explicitly to shared capture helpers. Keep the original gates, feature/stage names, viewports, waits, and recording options. Use distinct filenames for distinct stages and keep screenshots, reports, and video together.
  • Retain successful and failed evidence. Report actual allocated paths, including relocated filename overrides. Manually delete only exact owned directories after review; never clear shared parents before a replay. Disposable build/media fixtures and owned temporary raw video may keep their cleanup. New synthetic captures do not recover overwritten evidence.
  • Timeout diagnostics use fresh children beneath their existing diagnostic parent. Mantis retains every capture attempt under an invocation-owned directory and refuses to overwrite reports. Real-Gateway suites, chat-outbox-*, and chat-attachment-read-lifecycle remain separate owners; coordinate before claiming replay-safe retention there.
  • Treat recording as validation, not only demo capture. If the recorder fails or shows surprising behavior, stop, fix the behavior, add or update a regression test, then rerecord.
  • If visual proof is blocked, state the exact blocker and still report the textual E2E evidence.
Show full SKILL.md (342 more words)Show less

Mock Pattern

Start the app server, install the mock before page.goto, then assert both Gateway traffic and visible UI:

ts
const server = await startControlUiE2eServer();
const page = await context.newPage();
const gateway = await installMockGateway(page, {
  historyMessages: [{ role: "assistant", content: [{ type: "text", text: "Ready." }] }],
});

await page.goto(`${server.baseUrl}chat`);
await page.locator(".agent-chat__composer-combobox textarea").fill("hello");
await page.getByRole("button", { name: "Send message" }).click();

const request = await gateway.waitForRequest("chat.send");
await gateway.emitChatFinal({ runId: String(request.params.idempotencyKey), text: "Done." });
await page.getByText("Done.").waitFor();

Extend installMockGateway with typed scenario options or method responses when a new flow needs more Gateway surface.

Run Inspector evidence

Use withControlUiRunInspector from ui/src/test-helpers/control-ui-run-inspector.ts for collection. It owns a separate page, closes it on success or failure, and leaves the caller's Chat page and unsent draft in place. Use preparePage for a mock Gateway or the campaign's existing per-tab authentication setup; a shared browser context does not copy another tab's session-storage token. The helper uses RunInspectorSelector and activityRunInspectorSelectorHref from the rendered component's model, including a selected receipt's decision cursor.

The rendered panel's data-run-id and data-execution-id identify the returned present identity, not just the requested URL. The selected detail's data-receipt-selector-id is DecisionReceiptDisplayV1.selectorId. Missing or ambiguous identity and missing selected receipts must not be treated as matches. These decision selectors are separate from the optional terminal transcript key agent.wait.terminalReceipt.assistantTranscriptIdempotencyKey; never manufacture that key from a DOM selector or history row, or claim its absence is repaired by Inspector evidence. For sidebar run state, the current accessible label is Active run, not Running.

Standalone Recording

For narrated captions, eased target zooms, or fast-forwarded pauses, use the proof-video dev skill. Its standalone template records raw evidence plus a cue sidecar and renders a captioned MP4 with system ffmpeg; keep the raw capture and attach the inspected polished video.

When recording an already-running mocked Control UI URL, use a temporary Playwright script or playwright test spec and keep the recording flow focused:

  • Open the mock URL, interact through stable data-* selectors or user-facing role selectors, and wait on asserted states instead of relying on fixed sleeps.
  • Assert both visible UI state and mocked Gateway traffic for request-driven flows. For example, verify the expected count/row is visible and that sessions.list was called with the expected search, offset, and limit.
  • Use short sleeps only after assertions to make the captured video readable.
  • Store the generated video in the invocation's fresh allocated directory; do not commit it or remove older captures.

© openclaw, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

SKILL.md and 1 other file in .agents/skills/control-ui-e2e of openclaw/openclaw.

  • SKILL.md
  • agents/openai.yaml

Open the folder on GitHubat commit 1eb5970

Compare with similar skills

Control UI E2E next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Control UI E2E compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Control UI E2E this skillopenclaw/openclaw392k—~2.9kAutomated safety check: PassMIT
Web Testing with Playwright and Vitestwithkynam/vibecode-pro-max-kit1.1k—~892Automated safety check: PassApache-2.0
Test CommanderEliasOulkadi/shokunin114—~3kAutomated safety check: NotesMIT
Svelte Testingspences10/sveltest113—~579Automated safety check: PassMIT
Playwright Testingchongdashu/vibejam-starter-pack149—~2.1kAutomated safety check: PassNone
Ha Frontend Testinghome-assistant/frontend5.7k—~1.7kAutomated safety check: PassApache-2.0

Similar skills

  • Web Testing with Playwright and Vitest

    withkynam/vibecode-pro-max-kit

    Covers web testing from unit to E2E, load, visual, accessibility and security checks, with Playwright, Vitest and k6 guides plus a Playwright setup script.

    1.1k GitHub stars~892 tokensUpdated 3 mo ago
    Testing & QAAuto-check passed
  • Test Commander

    EliasOulkadi/shokunin

    Generate unit, integration, E2E, and visual regression tests following the Testing Trophy methodology (80% integration).

    114 GitHub stars~3k tokensUpdated 3 days ago
    Testing & QAAuto-check: notes
  • Svelte Testing

    spences10/sveltest

    Fix and create Svelte 5 tests with vitest-browser-svelte and Playwright.

    113 GitHub stars~579 tokensUpdated today
    Testing & QAAuto-check passed
  • Playwright Testing

    chongdashu/vibejam-starter-pack

    Plan, implement, and debug frontend tests: unit/integration/E2E/visual/a11y.

    149 GitHub stars~2.1k tokensUpdated 5 mo ago
    Testing & QAAuto-check passed
  • Ha Frontend Testing

    home-assistant/frontend

    Home Assistant frontend testing and validation workflow. An agent skill from home-assistant/frontend.

    5.7k GitHub stars~1.7k tokensUpdated today
    Testing & QAAuto-check passed
  • Playwright Testing

    chongdashu/vibejam-starter-pack

    Plan, implement, and debug frontend tests: unit/integration/E2E/visual/a11y.

    149 GitHub stars~2.2k tokensUpdated 5 mo ago
    Testing & QAAuto-check passed

More from openclaw/openclaw

All 93 skills in this repo
  • Openclaw Live Updater

    openclaw/openclaw

    Maintain the canonical live OpenClaw main checkout, macOS LaunchAgent-managed Gateway, local macOS app, exact-head main CI, and recurring full release validation.

    392k GitHub stars~3.7k tokensUpdated today
    Auto-check passed
  • Tmux

    openclaw/openclaw

    Control tmux sessions/panes for interactive CLIs: list, capture output, send keys, paste text, monitor prompts.

    392k GitHub starsUsed in 2 repos~640 tokens
    Auto-check passed
  • Feishu Doc

    openclaw/openclaw

    Feishu document read/write workflows. An agent skill from openclaw/openclaw.

    392k GitHub stars~516 tokensUpdated today
    Auto-check passed
  • Openclaw PR Maintainer

    openclaw/openclaw

    Review, triage, repair, or land OpenClaw issues and pull requests with current-source evidence and the native maintainer workflow.

    392k GitHub stars~2.2k tokensUpdated today
    Auto-check passed
  • Browser Automation

    openclaw/openclaw

    A skill your agent uses when controlling web pages with the OpenClaw browser tool, especially multi-step flows, login checks, tab management, or recovery from stale refs/timeouts.

    392k GitHub stars~2.9k tokensUpdated today
    Auto-check passed
  • Clawsweeper

    openclaw/openclaw

    A skill your agent uses for all ClawSweeper work: OpenClaw issue/PR sweep reports, repair jobs, cloud fix PRs, @clawsweeper maintainer mention commands, trusted ClawSweeper-reviewed…

    392k GitHub stars~3k tokensUpdated today
    Auto-check passed

Categories

Questions about Control UI E2E

What does Control UI E2E do?

A skill your agent uses when designing, testing, fixing, or extending the OpenClaw Control UI GUI, including UI stress-test galleries with feedback inputs, Vitest + Playwright end-to-end checks…. Control UI E2E is an agent skill from openclaw/openclaw. Use when designing, testing, fixing, or extending the OpenClaw Control UI GUI, including UI stress-test galleries with feedback inputs, Vitest + Playwright end-to-end checks, mocked Gateway flows, screenshots/videos, or agent-verifiable browser proof.

When should I use Control UI E2E?

Control UI E2E fits situations like: extending the OpenClaw Control UI GUI; including UI stress-test galleries with feedback inputs; vitest + Playwright end-to-end checks; mocked Gateway flows.

How do I install Control UI E2E in Claude Code?

Run `npx skills add openclaw/openclaw --skill control-ui-e2e -a claude-code`. Or copy the skill folder (.agents/skills/control-ui-e2e in openclaw/openclaw) into .claude/skills/control-ui-e2e in your project. Claude Code loads it when a task matches its description.

How do I install Control UI E2E in Codex?

Run `npx skills add openclaw/openclaw --skill control-ui-e2e -a codex`. Or copy the skill folder (.agents/skills/control-ui-e2e in openclaw/openclaw) into .agents/skills/control-ui-e2e in your project. Codex loads it when a task matches its description.

Can I use Control UI E2E in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add openclaw/openclaw --skill control-ui-e2e -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/control-ui-e2e, .gemini/skills/control-ui-e2e, .github/skills/control-ui-e2e and .opencode/skills/control-ui-e2e in your project.

What does Control UI E2E need to run?

Going by SKILL.md and its folder, Control UI E2E needs the command-line tools its instructions call (pnpm, node and playwright).

Does Control UI E2E access the network?

SKILL.md names 1 domain. As links in the text: docs.openclaw.ai. This is read from the text; nothing was executed.

Is Control UI E2E safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Control UI E2E use?

Control UI E2E is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Control UI E2E use?

About 2.9k tokens (SKILL.md is roughly 11k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Control UI E2E?

Skills that share tags, products or a category with Control UI E2E: Web Testing with Playwright and Vitest (withkynam/vibecode-pro-max-kit, 1.1k stars), Test Commander (EliasOulkadi/shokunin, 114 stars), Svelte Testing (spences10/sveltest, 113 stars) and Playwright Testing (chongdashu/vibejam-starter-pack, 149 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Control UI E2E?

openclaw (a GitHub organization) maintains it in openclaw/openclaw, which has 391,610 GitHub stars. The repository holds 93 skills in this directory. The repository was last updated on October 8, 2026.

Source: openclaw/openclaw on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.