Agent skill

cmux Testing and Verification

by manaflow-ai in manaflow-ai/cmux

Chooses scoped cmux verification checks, adds behavioral and regression tests, and validates Swift test targets and wiring, with local versus CI guidance.

Custom licenceAuto-check passedTesting & QA

Install cmux Testing and Verification

skills CLI
$ npx skills add manaflow-ai/cmux --skill cmux-testing -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install manaflow-ai/cmux cmux-testing --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/manaflow-ai/cmux.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/cmux-testing .claude/skills/cmux-testing && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
cmux-testing
GitHub stars
28k
Token cost
~1.4k tokens
SKILL.md length
681 words
Files
10 (incl. references)
Skills in repo
22
Repo updated
First seen
Licence
Custom licence

At a glance

Chooses scoped cmux verification checks, adds behavioral and regression tests, and validates Swift test targets and wiring, with local versus CI guidance.

  • Deciding which local or CI checks a cmux change needs
  • SKILL.md covers Choose the first check, Reproduce and repair, Test wiring and Test quality, plus 3 more sections
  • Calls python3
  • Adding a behavioral regression test for a reported cmux bug

What it does

The skill starts by choosing the cheapest useful check. Commands run through `scripts/verify-local.py`: the default picks checks and parses changed Swift, `--all` runs the full CI static recipe, `--only swift-syntax --swift-changed` parses current edits, and `--only test-wiring` checks new Swift test-file wiring. Parsing checks syntax only and does not typecheck or run tests. Repository commands should run only from a trusted checkout, since even `--help` loads repository code.

For bugs, it keeps a focused command that fails on the reported symptom, then makes two commits: a failing behavioral regression first, then the fix, running the same command on both and recording commit SHAs and results. Setup failures and zero executed tests do not count as proof. Other helpers include `scripts/ui-test` for one frame per action, `scripts/ui-lab/ui-lab.py` for rendering view code to PNGs in light and dark, and `scripts/run-e2e.sh` with a JSON tour for dogfooding the app from CI. References cover local versus CI validation, Swift Testing migration and remote tmux sizing.

When your agent uses it

  • Deciding which local or CI checks a cmux change needs
  • Adding a behavioral regression test for a reported cmux bug
  • Checking that a new Swift test file is wired into its test target
  • Previewing view code changes as light and dark PNGs without a full build

Example prompts

  • “Which verification should I run for this change to the sidebar Swift code?”
  • “Add a failing regression test for the tab-switching bug first, then fix it.”
  • “Check that my new Swift test file is wired into the test target.”
  • “Render the settings view in light and dark with ui-lab.”

Requirements

  • A trusted checkout of the cmux repository
  • Python 3 for `scripts/verify-local.py`

What it can do on your machine

Read from SKILL.md and the folder at commit f6532b1. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Shell commands in SKILL.md call:

    • python3

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

cmux Testing and Verification loads about 1.4k tokens when it runs, and up to ~9.3k if it reads all its reference files. Until then it costs about 46 tokens; SKILL.md has 681 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~46
When it runs · the whole SKILL.md, loaded when a task matches
~1.4k
With references · SKILL.md plus every file in references/, read only if the agent opens them
~9.3k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

Its licence (Custom licence) doesn't allow us to republish the file, so here is its outline and opening line. It has 681 words (~1,400 tokens).

“Run repository commands only from a trusted checkout; even verify-local.py --help and --list load repository code.”

— opening of SKILL.md by manaflow-ai, Custom licence
name
cmux-testing

Read the full SKILL.md on GitHub

Files

SKILL.md and 9 other files (references) in skills/cmux-testing of manaflow-ai/cmux.

  • SKILL.md
  • agents/openai.yaml
  • references/dogfood-scenarios.md
  • references/local-vs-ci-validation.md
  • references/pr-ci-coverage.md
  • references/regression-and-quality.md
  • references/remote-tmux-sizing-e2e.md
  • references/swift-testing-migration.md
  • references/ui-lab.md
  • references/ui-test-frames.md

Open the folder on GitHubat commit f6532b1

Compare with similar skills

cmux Testing and Verification next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

cmux Testing and Verification compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
cmux Testing and Verification this skillmanaflow-ai/cmux28k—~1.4kAutomated safety check: PassCustom licence
Senior QAnicepkg/auto-company1943 repos~1.1kAutomated safety check: NotesNone
Designing TestsCloudAI-X/claude-workflow-v21.4k1 repos~1.5kAutomated safety check: PassMIT
Test Pyramidkubernetes-sigs/agent-sandbox4.2k—~1.6kAutomated safety check: PassApache-2.0
Studio Testingsupabase/supabase111k—~2.2kAutomated safety check: PassApache-2.0
QAdutradotdev/quokka108—~1kAutomated safety check: PassMIT

Similar skills

  • Senior QA

    nicepkg/auto-company

    Comprehensive QA and testing skill for quality assurance, test automation, and testing strategies for ReactJS, NextJS, NodeJS applications.

    194 GitHub starsUsed in 3 repos~1.1k tokens
    Testing & QAAuto-check: notes
  • Designing Tests

    CloudAI-X/claude-workflow-v2

    Designs and implements testing strategies for any codebase. An agent skill from CloudAI-X/claude-workflow-v2.

    1.4k GitHub starsUsed in 1 repo~1.5k tokens
    Testing & QAAuto-check passed
  • Test Pyramid

    kubernetes-sigs/agent-sandbox

    Official

    Analyze the repo's unit and E2E tests and propose rebalancing toward a test pyramid — which E2E tests (or assertions inside them) can be covered by unit tests, which unit-level gaps genuinely need…

    4.2k GitHub stars~1.6k tokensUpdated today
    Testing & QAAuto-check passed
  • Studio Testing

    supabase/supabase

    Official

    Testing strategy for Supabase Studio. An agent skill from supabase/supabase.

    111k GitHub stars~2.2k tokensUpdated today
    Testing & QAAuto-check passed
  • QA

    dutradotdev/quokka

    Run quokka's real-device E2E layer (QA strategy layer 5): drive the real qk binary against connected iPhone/Android devices through tmux, apply deterministic verifiers, and write a report.

    108 GitHub stars~1k tokensUpdated 2 mo ago
    Testing & QAAuto-check passed
  • Finds out why a passing test suite sits on top of a broken app by comparing what the running app does with what the tests claim, using Reticle.

    1.2k GitHub stars~1.1k tokensUpdated today
    Testing & QAAuto-check passed

More from manaflow-ai/cmux

All 22 skills in this repo
  • cmux Diagnostics

    manaflow-ai/cmux

    Runs a read-only health check for cmux and explains what it finds, covering the CLI and socket, settings, agent hooks, session restore and notifications.

    28k GitHub starsUsed in 1 repo~819 tokens
    Auto-check passed
  • cmux Settings Editor

    manaflow-ai/cmux

    Reads and changes cmux preferences in `~/.config/cmux/cmux.json` through a helper that validates each change before writing, covering terminal, browser, viewers and shortcuts.

    28k GitHub stars~2.3k tokensUpdated today
    Auto-check passed
  • cmux Backend Rules

    manaflow-ai/cmux

    Sets the backend TypeScript and Cloud VM rules for cmux: Effect-based services, thin route handlers, Postgres as source of truth, migrations and provider secrets.

    28k GitHub starsUsed in 1 repo~682 tokens
    Auto-check passed
  • Cmux Debugging Guide

    manaflow-ai/cmux

    Covers debug logging, the Debug menu, profiling rules and runtime pitfalls for working on the cmux macOS terminal app.

    28k GitHub starsUsed in 1 repo~1.1k tokens
    Auto-check passed
  • Workflow rules for the Ghostty submodule in cmux: rebuilding GhosttyKit.xcframework, pushing fork changes and updating the parent submodule pointer safely.

    28k GitHub starsUsed in 1 repo~580 tokens
    Auto-check passed
  • Cmux Release Workflow

    manaflow-ai/cmux

    Runs the cmux release process: gathers per-PR changelog lines since the last tag, bumps the version, guards against a bad tag, then tags and pushes the macOS DMG release.

    28k GitHub starsUsed in 1 repo~856 tokens
    Auto-check passed

Categories

Questions about cmux Testing and Verification

What does cmux Testing and Verification do?

Chooses scoped cmux verification checks, adds behavioral and regression tests, and validates Swift test targets and wiring, with local versus CI guidance. The skill starts by choosing the cheapest useful check.py`: the default picks checks and parses changed Swift, `--all` runs the full CI static recipe, `--only swift-syntax --swift-changed` parses current edits, and `--only test-wiring` checks new Swift test-file wiring.

When should I use cmux Testing and Verification?

cmux Testing and Verification fits situations like: deciding which local or CI checks a cmux change needs; adding a behavioral regression test for a reported cmux bug; checking that a new Swift test file is wired into its test target; previewing view code changes as light and dark PNGs without a full build.

How do I install cmux Testing and Verification in Claude Code?

Run `npx skills add manaflow-ai/cmux --skill cmux-testing -a claude-code`. Or copy the skill folder (skills/cmux-testing in manaflow-ai/cmux) into .claude/skills/cmux-testing in your project. Claude Code loads it when a task matches its description.

How do I install cmux Testing and Verification in Codex?

Run `npx skills add manaflow-ai/cmux --skill cmux-testing -a codex`. Or copy the skill folder (skills/cmux-testing in manaflow-ai/cmux) into .agents/skills/cmux-testing in your project. Codex loads it when a task matches its description.

Can I use cmux Testing and Verification in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add manaflow-ai/cmux --skill cmux-testing -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/cmux-testing, .gemini/skills/cmux-testing, .github/skills/cmux-testing and .opencode/skills/cmux-testing in your project.

What does cmux Testing and Verification need to run?

Going by SKILL.md and its folder, cmux Testing and Verification needs the command-line tools its instructions call (python3). Our summary lists: A trusted checkout of the cmux repository; Python 3 for `scripts/verify-local.py`.

Does cmux Testing and Verification access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is cmux Testing and Verification safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does cmux Testing and Verification use?

cmux Testing and Verification has a licence file (the repository's licence) that doesn't match a standard licence. Read it on GitHub before reusing the skill.

How many tokens does cmux Testing and Verification use?

About 1.4k tokens (SKILL.md is roughly 5.6k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full. Its references folder adds about 7.9k tokens, read only when the agent opens those files.

What are the alternatives to cmux Testing and Verification?

Skills that share tags, products or a category with cmux Testing and Verification: Senior QA (nicepkg/auto-company, 194 stars), Designing Tests (CloudAI-X/claude-workflow-v2, 1.4k stars), Test Pyramid (kubernetes-sigs/agent-sandbox, 4.2k stars) and Studio Testing (supabase/supabase, 111k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains cmux Testing and Verification?

manaflow-ai (a GitHub organization) maintains it in manaflow-ai/cmux, which has 28,044 GitHub stars. The repository holds 22 skills in this directory. The repository was last updated on October 9, 2026.

Source: manaflow-ai/cmux on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.