Official agent skill

Test Electron App

by PostHog in PostHog/code

Drive the real running PostHog Electron app (live tRPC, workspace-server, real data) over CDP with agent-browser.

OfficialMITAuto-check passedTesting & QA

Install Test Electron App

skills CLI
$ npx skills add PostHog/code --skill test-electron-app -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install PostHog/code test-electron-app --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/PostHog/code.git skills-src && mkdir -p .claude/skills && cp -r skills-src/.claude/skills/test-electron-app .claude/skills/test-electron-app && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
test-electron-app
GitHub stars
179
Token cost
~2.9k tokens
SKILL.md length
1,333 words
Files
1
Skills in repo
10
Repo updated
First seen
Licence
MIT

At a glance

Drive the real running PostHog Electron app (live tRPC, workspace-server, real data) over CDP with agent-browser.

  • Works in 3 steps: Save to and report an absolute path… → Offer to open it for them, and on a yes… → The user can also open it themselves at…
  • Interact with the running app
  • SKILL.md covers Prerequisites, Load the canonical commands, The loop and Screenshots, plus 3 more sections
  • Calls pnpm, curl and npm

What it does

Test Electron App is an agent skill from PostHog/code, published by the product's own GitHub organization. Drive the real running PostHog Electron app (live tRPC, workspace-server, real data) over CDP with agent-browser. Connect to the running app on port 9222, snapshot the accessibility tree to verify changes, click/type/navigate, and screenshot the actual desktop app only when explicitly asked. Use when asked to test, verify, dogfood, screenshot or interact with the running app. For regression specs use the Playwright E2E suite.

Its SKILL.md is about 2.9k tokens, which your agent loads only when the skill is triggered. It is a single SKILL.md file with no bundled scripts.

It sits in Testing & QA, covering Browser automation, gRPC and Protobuf and End-to-end testing. It works with Electron, PostHog, Playwright and tRPC. The repository describes itself as: The era of self-driving development is here. The licence is MIT.

When your agent uses it

  • Interact with the running app
  • Tasks that involve Browser automation
  • Tasks that involve gRPC and Protobuf

Example prompts

  • “/test-electron-app”

Requirements

  • Node.js
  • Pre-approved tools (allowed-tools): Bash(agent-browser:*), Bash(npx agent-browser:*), Bash(pnpm app:cdp*)

Workflow steps

3 steps, taken from the first numbered list in SKILL.md.

  1. Save to and report an absolute path (/tmp/app.png, never a bare out.png). Claude Code renders absolute paths as clickable links, so the…
  2. Offer to open it for them, and on a yes run open /tmp/app.png (macOS opens it in Preview). If the request was clearly "screenshot it so I…
  3. The user can also open it themselves at any time with !open /tmp/app.png.

What it can do on your machine

Read from SKILL.md and the folder at commit a4c32df. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves these tools, so the agent can use them without asking each time:

    • Bash(agent-browser:*)
    • Bash(npx agent-browser:*)
    • Bash(pnpm app:cdp*)

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Shell commands in SKILL.md call:

    • pnpm
    • curl
    • npm

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    Links to these hosts (documentation or services it may open):

    • github.com

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Test Electron App loads about 2.9k tokens when it runs. Until then it costs about 112 tokens; SKILL.md has 1,333 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~112
When it runs · the whole SKILL.md, loaded when a task matches
~2.9k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from PostHog/code at commit a4c32df, republished under its MIT licence (© PostHog). 1,333 words, ~2,881 tokens.

Download SKILL.mdSave it as .claude/skills/test-electron-app/SKILL.md (or your agent's skills folder).
name
test-electron-app
description
Drive the real running PostHog Electron app (live tRPC, workspace-server, real data) over CDP with agent-browser. Connect to the running app on port 9222, snapshot the accessibility tree to verify changes, click/type/navigate, and screenshot the actual desktop app only when explicitly asked. Use when asked to test, verify, dogfood, screenshot or interact with the running app. For regression specs use the Playwright E2E suite.
allowed-tools
Bash(agent-browser:*), Bash(npx agent-browser:*), Bash(pnpm app:cdp*)

Test the real PostHog Electron app

Drive the actual running app over the Chrome DevTools Protocol with agent-browser. The dev app already launches with --remote-debugging-port=9222 (see apps/code/package.json), so an agent can connect, snapshot the UI, interact, and (only when asked) screenshot the live app.

This exercises real state: live tRPC, workspace-server, GitHub/Slack and whatever profile is signed into ~/.posthog-code. Pick the right surface:

GoalTool
Verify or screenshot a change in the real app, live datathis skill (agent-browser + CDP :9222)
Regression coverage in CIPlaywright E2E (apps/code/tests/e2e/)

Prerequisites

bash
npm i -g agent-browser && agent-browser install   # once

The app must be running with remote debugging. pnpm dev (or pnpm dev:code) already passes --remote-debugging-port=9222. Preflight + connect:

bash
pnpm app:cdp        # checks agent-browser + that the app is up on :9222, then connects

If it reports the app is not reachable, the app isn't running. Start it, then retry pnpm app:cdp.

Launching the app yourself (background / no TTY)

Do not use pnpm dev from a non-interactive shell. It builds deps then hands off to the phrocs TUI process-multiplexer, which aborts without a controlling terminal (bubbletea: could not open TTY: /dev/tty: device not configured). An agent in a background shell has no TTY, so pnpm dev cannot launch it headlessly.

Build the workspace deps (TTY-safe, the same step pnpm dev runs first), then launch only the Electron app with its stdin held open:

bash
pnpm build:deps                        # turbo build of @posthog/code deps
tail -f /dev/null | pnpm dev:code      # run in the background; leave it running

pnpm dev:code is pnpm --filter code start, i.e. electron-vite dev --watch — just the app, no phrocs TUI; it watches and rebuilds main/preload and hot-reloads the renderer (fine for screenshotting/interacting). The CDP port (:9222) is opened by the app itself in dev (the remote-debugging-port switch in apps/code/src/main/bootstrap.ts), not by a CLI flag.

The tail -f /dev/null | prefix is a harmless guard and no longer strictly required. The old electron-forge start ran an interactive "type rs to restart" stdin reader that hit EOF in a no-stdin shell, treated it as quit, and tore the Electron window down before the CDP port ever opened. electron-vite dev has no such reader, so a backgrounded pnpm dev:code with no stdin stays up on its own; keeping the pipe does no harm.

Then wait for the port and connect (poll, don't sleep blindly):

bash
until curl -s localhost:9222/json/version >/dev/null; do sleep 1; done
agent-browser connect 9222
After you launch it: leave it up, then idle-shutdown

Never auto-kill an instance you didn't start. If pnpm app:cdp found the app already up on :9222, it's the user's own — when you're done just agent-browser close your session and leave the app running.

When you launched it, don't tear it down the instant you finish. The app survives between turns, so leave it up and end your turn by telling the user it's still running and asking if they want anything else — a follow-up needs no relaunch.

So a forgotten app doesn't linger, arm a 10-minute idle watchdog at launch. Touch a marker file on launch and after every interaction; the watchdog tears the app down once that file sits untouched for 10 minutes (re-touching resets the clock):

bash
touch /tmp/posthog-dev-lastuse        # arm now; re-run after each interaction

Then start the watchdog once, as a background task (touch the marker first or it fires immediately). It polls, then self-exits after it fires or once the marker is removed:

bash
while sleep 30; do
  last=$(stat -f %m /tmp/posthog-dev-lastuse 2>/dev/null || echo 0)
  [ $(( $(date +%s) - last )) -ge 600 ] && break
done
pid=$(pgrep -f "remote-debugging-port=9222" | head -1)
[ -n "$pid" ] && kill -TERM "-$(ps -o pgid= -p "$pid" | tr -d ' ')" 2>/dev/null

To stop early (user says "done" / "shut it down"): close the session, group-kill the app, and drop the marker so the watchdog exits:

bash
agent-browser close
pid=$(pgrep -f "remote-debugging-port=9222" | head -1)
[ -n "$pid" ] && kill -TERM "-$(ps -o pgid= -p "$pid" | tr -d ' ')" 2>/dev/null
rm -f /tmp/posthog-dev-lastuse

Group-killing the launcher (kill -TERM -PGID) takes down tail, pnpm, electron-vite and Electron together, so nothing — not even the tail -f /dev/null stdin pipe — lingers. Matching remote-debugging-port=9222 hits only your dev instance (prod has no debug port and a separate posthog-code-dev profile), so it never touches the user's app. Verify with curl -s localhost:9222/json/version (fails) and pgrep -fl posthog-code-dev (empty).

Load the canonical commands

agent-browser serves version-matched docs. Read them before driving:

bash
agent-browser skills get electron     # Electron-over-CDP workflow (authoritative)
agent-browser skills get core         # snapshot/interact/screenshot reference

The loop

bash
agent-browser connect 9222                      # attach (skip if you ran pnpm app:cdp)
agent-browser snapshot -i                       # interactive elements only (the app is already dark)
agent-browser click @e5                          # act on a ref from the snapshot
agent-browser snapshot -i                        # ALWAYS re-snapshot after the UI changes — this is how you verify
agent-browser close                              # done; free the session

Verifying a change is snapshot, not screenshot. The accessibility tree tells you what is on screen for almost no tokens, and it is how you confirm a test worked. Do not capture a screenshot to check your own work — only run screenshot when the user explicitly asks to see the app (see Screenshots).

Refs (@e1, @e2, …) are reassigned on every snapshot and go stale the moment the UI changes. Re-snapshot before the next ref interaction.

The renderer uses data-testid heavily, so prefer stable locators over refs when you know the target:

bash
agent-browser find testid new-task-button click
agent-browser find role button click --name "New task"
agent-browser find text "Settings" click

Screenshots

Only screenshot when the user explicitly asks for one ("screenshot", "show me", "what does it look like"). To confirm a change worked, use agent-browser snapshot — the accessibility tree is the cheap default and is almost always enough. Auto-capturing a screenshot to "confirm" your work just burns image tokens.

bash
agent-browser screenshot /tmp/app.png                  # viewport (absolute path = clickable)
agent-browser screenshot --full /tmp/app.png           # full page instead of viewport

Navigate to the target view first (click through the UI), then capture. agent-browser prints the saved path. Repeated captures reuse the connected session, so batches are fast.

Show full SKILL.md (568 more words)Show less
Always let the user open the capture

When the user asked for a screenshot, they want to look at it, so close the loop every time:

  1. Save to and report an absolute path (/tmp/app.png, never a bare out.png). Claude Code renders absolute paths as clickable links, so the path itself opens the PNG on click.
  2. Offer to open it for them, and on a yes run open /tmp/app.png (macOS opens it in Preview). If the request was clearly "screenshot it so I can see it", just open it instead of asking.
  3. The user can also open it themselves at any time with !open /tmp/app.png.

Repo specifics

  • Port: 9222 (override with POSTHOG_CODE_CDP_PORT). Collides with Chrome's default debugging port; if connect attaches to the wrong target, list and pick the PostHog window: agent-browser tab then agent-browser tab --url "*".
  • Multiple targets: the app has a main renderer window (page title contains "PostHog") plus possible webviews/devtools. agent-browser tab lists them; switch with agent-browser tab <index>.
  • Never pass --color-scheme dark: that global flag makes agent-browser apply device emulation that forces a 1280x720 viewport and renders this Electron window blank, and it sticks in the daemon (only restarting the agent-browser daemon clears it, not close). The app is already dark, so plain agent-browser snapshot / agent-browser screenshot is what you want.
  • Auth / data: you drive whatever is signed into ~/.posthog-code. If the app shows onboarding or sign-in, that is the real boot state. Do not mutate production data (don't create real tasks/PRs) while exploring.
  • Boot timing: after launching the app, give it a few seconds before connecting; the renderer settles after #root > * appears and "Loading" clears.

Running alongside prod

PostHog orchestrates the agent, so the usual loop is: prod (the installed app) runs the agent, and the dev build (pnpm dev) is the system under test. They coexist by design (apps/code/src/main/bootstrap.ts): dev runs as posthog-code-dev with its own app name, userData and single-instance lock, so it never collides with prod.

  • agent-browser always targets dev. Only the dev build exposes CDP on :9222; prod has no debug port, so connect 9222 can't accidentally drive prod.
  • Separate auth/state. The dev instance has its own posthog-code-dev profile; it is not signed in just because prod is. Sign into the dev window once; its state persists.
  • One dev instance only. Dev's single-instance lock, fixed dev callback port (8238) and :9222 mean a second pnpm dev collides and quits. Run prod + one dev.
  • What reloads. Renderer/UI changes hot-reload; just re-snapshot. Main-process/Electron changes need a dev restart to take effect.

Troubleshooting

  • Connection refused on :9222: the app isn't running with the debug flag. Start it (see Launching the app yourself for the headless recipe). Verify the port: lsof -i :9222 or curl -s localhost:9222/json/version.
  • App launches then immediately exits; CDP never opens: the old electron-forge "quit on stdin EOF" behavior is gone under electron-vite, so the tail -f /dev/null | prefix is optional and not the fix here. This is now usually a build error, the single-instance lock (another dev instance running) or a crash — check the pnpm dev:code output. Note pnpm dev cannot run headlessly at all — its phrocs TUI needs a TTY; use pnpm dev:code.
  • Snapshot is empty / wrong window: you're on the wrong target. Run agent-browser tab and switch to the "PostHog" page.
  • Can't type into an input: try agent-browser keyboard type "text" (types at current focus) or agent-browser keyboard inserttext "text" to bypass key events.

© PostHog, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

Just SKILL.md in .claude/skills/test-electron-app of PostHog/code.

Open the folder on GitHubat commit a4c32df

Compare with similar skills

Test Electron App next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Test Electron App compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Test Electron App this skillPostHog/code179—~2.9kAutomated safety check: PassMIT
Test Electron AppPostHog/posthog-foss721—~3kAutomated safety check: PassMIT
playwright-cli Browser Automationgithub/gh-aw5.3k23 repos~2.8kAutomated safety check: PassMIT
E2E Agent Browserjh941213/my-cc-harness126—~3.1kAutomated safety check: NotesNone
UI Previewtingly-dev/tingly-box351—~2kAutomated safety check: NotesMPL-2.0
E2E VerificationChorus-AIDLC/Chorus1.2k—~1.5kAutomated safety check: NotesAGPL-3.0

Similar skills

  • Test Electron App

    PostHog/posthog-foss

    Official

    Drive the real running PostHog Electron app (live tRPC, workspace-server, real data) over CDP with agent-browser.

    721 GitHub stars~3k tokensUpdated today
    Testing & QAAuto-check passed
  • Official

    Drives a real browser from the command line with playwright-cli to open pages, interact, mock requests, save state and work with Playwright tests.

    5.3k GitHub starsUsed in 23 repos~2.8k tokens
    Testing & QAAuto-check passed
  • E2E Agent Browser

    jh941213/my-cc-harness

    E2E test automation using agent-browser CLI. An agent skill from jh941213/my-cc-harness.

    126 GitHub stars~3.1k tokensUpdated 2 mo ago
    Testing & QAAuto-check: notes
  • UI Preview

    tingly-dev/tingly-box

    Capture headless-Chrome screenshots of the tingly-box frontend (running locally in mock mode) so frontend changes can be visually verified in environments without a real browser.

    351 GitHub stars~2k tokensUpdated today
    Testing & QAAuto-check: notes
  • E2E Verification

    Chorus-AIDLC/Chorus

    A skill your agent uses when manually verifying a Chorus frontend change in a real browser — finding local login credentials, driving the running dev server with the Playwright MCP, logging in…

    1.2k GitHub stars~1.5k tokensUpdated today
    Testing & QAAuto-check: notes
  • E2E Playwright

    Asvarox/allkaraoke

    Run, write, and debug Playwright E2E tests for this project.

    261 GitHub stars~876 tokensUpdated today
    Testing & QAAuto-check passed

More from PostHog/code

All 10 skills in this repo
  • Canvas Templates

    PostHog/code

    Official

    How PostHog "canvas" dashboards work end-to-end — the two rendering tiers (json-render vs freeform React-in-iframe), the agent system prompts that steer each, and the RIGHT way to fetch PostHog data…

    179 GitHub stars~2.2k tokensUpdated 2 mo ago
    Auto-check passed
  • Dynamic Workflows

    PostHog/code

    Official

    How to write JavaScript workflow scripts for the workflow tool - fanning work out across many isolated subagents with agent(), parallel(), and pipeline(), then synthesizing one result.

    179 GitHub stars~2.3k tokensUpdated 2 mo ago
    Auto-check passed
  • MCP Servers

    PostHog/code

    Official

    Install, configure, authenticate, and troubleshoot MCP (Model Context Protocol) servers for this agent.

    179 GitHub stars~1.9k tokensUpdated 2 mo ago
    Auto-check passed
  • Merging PRs

    PostHog/code

    Official

    Merge a PR into main through the Trunk merge queue and babysit it until it lands.

    179 GitHub stars~1.3k tokensUpdated 2 mo ago
    Auto-check passed
  • Onboarding Videos

    PostHog/code

    Official

    Add, replace, and optimize the looping demo videos in the onboarding "welcome" bento grid (packages/ui/src/features/onboarding).

    179 GitHub stars~1.7k tokensUpdated 2 mo ago
    Auto-check passed
  • Quill Code

    PostHog/code

    Official

    Edit the @posthog/quill design system locally and consume the change in this repo (posthog-code) before it is published to npm.

    179 GitHub stars~1.1k tokensUpdated 2 mo ago
    Auto-check passed

Questions about Test Electron App

What does Test Electron App do?

Drive the real running PostHog Electron app (live tRPC, workspace-server, real data) over CDP with agent-browser. Test Electron App is an agent skill from PostHog/code, published by the product's own GitHub organization. Drive the real running PostHog Electron app (live tRPC, workspace-server, real data) over CDP with agent-browser.

When should I use Test Electron App?

Test Electron App fits situations like: interact with the running app; tasks that involve Browser automation; tasks that involve gRPC and Protobuf.

How do I install Test Electron App in Claude Code?

Run `npx skills add PostHog/code --skill test-electron-app -a claude-code`. Or copy the skill folder (.claude/skills/test-electron-app in PostHog/code) into .claude/skills/test-electron-app in your project. Claude Code loads it when a task matches its description.

How do I install Test Electron App in Codex?

Run `npx skills add PostHog/code --skill test-electron-app -a codex`. Or copy the skill folder (.claude/skills/test-electron-app in PostHog/code) into .agents/skills/test-electron-app in your project. Codex loads it when a task matches its description.

Can I use Test Electron App in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add PostHog/code --skill test-electron-app -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/test-electron-app, .gemini/skills/test-electron-app, .github/skills/test-electron-app and .opencode/skills/test-electron-app in your project.

What does Test Electron App need to run?

Going by SKILL.md and its folder, Test Electron App needs the command-line tools its instructions call (pnpm, curl and npm). Our summary lists: Node.js. Its frontmatter pre-approves these tools: Bash(agent-browser:*), Bash(npx agent-browser:*), Bash(pnpm app:cdp*).

Does Test Electron App access the network?

SKILL.md names 1 domain. As links in the text: github.com. This is read from the text; nothing was executed.

Is Test Electron App safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Test Electron App use?

Test Electron App is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Test Electron App use?

About 2.9k tokens (SKILL.md is roughly 12k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Test Electron App?

Skills that share tags, products or a category with Test Electron App: Test Electron App (PostHog/posthog-foss, 721 stars), playwright-cli Browser Automation (github/gh-aw, 5.3k stars), E2E Agent Browser (jh941213/my-cc-harness, 126 stars) and UI Preview (tingly-dev/tingly-box, 351 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Test Electron App?

PostHog (a GitHub organization, an official publisher) maintains it in PostHog/code, which has 179 GitHub stars. The repository holds 10 skills in this directory. The repository was last updated on August 5, 2026.

Source: PostHog/code on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.