Agent skill

Test Examples

by mrchantey in mrchantey/beet

Run the workspace example smoke set to catch regressions unit tests miss, covering every feature gate combination with self-terminating examples and curl-probed servers.

Apache-2.0Auto-check passedTesting & QA

Install Test Examples

skills CLI
$ npx skills add mrchantey/beet --skill test-examples -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install mrchantey/beet test-examples --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/mrchantey/beet.git skills-src && mkdir -p .claude/skills && cp -r skills-src/.agents/skills/test-examples .claude/skills/test-examples && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
test-examples
GitHub stars
135
Token cost
~2.9k tokens
SKILL.md length
982 words
Files
1
Skills in repo
11
Repo updated
First seen
Licence
Apache-2.0

At a glance

Run the workspace example smoke set to catch regressions unit tests miss, covering every feature gate combination with self-terminating examples and curl-probed servers.

  • Works in 8 steps: Action (--features=action) → Scripting (--features=quickjs) → Router (--features=router,markdown /… → …
  • Verify the examples
  • SKILL.md covers Context, Smoke Set, Not Verifiable Via CLI (skip) and Instructions, plus 1 more section
  • Calls cargo and curl; needs OPENAI_API_KEY

What it does

Test Examples is an agent skill from mrchantey/beet. Run the workspace example smoke set to catch regressions unit tests miss, covering every feature gate combination with self-terminating examples and curl-probed servers. Use when asked to run or verify the examples.

Its SKILL.md is about 2.9k tokens, which your agent loads only when the skill is triggered. It is a single SKILL.md file with no bundled scripts.

It sits in Testing & QA, covering Unit testing. The repository describes itself as: A malleable engine built on interoperable standards. The licence is Apache-2.0.

When your agent uses it

  • Verify the examples
  • Tasks that involve Unit testing

Example prompts

  • “/test-examples”

Requirements

  • A credential in OPENAI_API_KEY

Workflow steps

8 steps, taken from the step headings in SKILL.md.

  1. Action (--features=action)
  2. Scripting (--features=quickjs)
  3. Router (--features=router,markdown / plus extras)
  4. Todo (--features=router,json)
  5. Net (--features=net,ureq,native-tls / --features=http_server)
  6. Per-crate examples
  7. Workspace ML (--features=examples,ml)
  8. BSX scenes (beet --main=.bsx)

What it can do on your machine

Read from SKILL.md and the folder at commit 38648b8. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Shell commands in SKILL.md call:

    • cargo
    • curl

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md. Its commands use curl, which can reach the network depending on how they are called.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names these keys or tokens, usually read from environment variables:

    • OPENAI_API_KEY

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Test Examples loads about 2.9k tokens when it runs. Until then it costs about 57 tokens; SKILL.md has 982 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~57
When it runs · the whole SKILL.md, loaded when a task matches
~2.9k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from mrchantey/beet at commit 38648b8, republished under its Apache-2.0 licence (© mrchantey). 982 words, ~2,877 tokens.

Download SKILL.mdSave it as .claude/skills/test-examples/SKILL.md (or your agent's skills folder).
name
test-examples
description
Run the workspace example smoke set to catch regressions unit tests miss, covering every feature gate combination with self-terminating examples and curl-probed servers. Use when asked to run or verify the examples.

Run Examples

Use the workspace examples as smoke tests to catch regressions that unit and integration tests may have missed.

Context

Not every example is verifiable from the CLI — many are 3D/2D Bevy windowed apps, browser-driven, or TUIs that block on stdin. This skill only exercises the ones that exit on their own (or that we can probe with curl while they run in the background).

Stream the entire output of each example into .agents/tmp/scratch.txt (overwrite for the first command, append for subsequent ones), then grep that file. This avoids reruns when checking multiple things.

Treat compile failures the same as the test-run skill: retry once, and if a mold linker error persists (RUST_MIN_STACK, "section sizes" etc) bump the workspace version = "0.0.9-dev.N" in Cargo.toml.

Long-running examples must always be wrapped in timeout (default 60s — drop it lower once you know the example exits faster). Server examples should be launched with run_in_background and killed once the probe has succeeded.

Smoke Set

The set below was chosen so that each feature gate combination is exercised by at least one example, and each crate has at least one self-terminating verifier. Most examples just need an OK exit status; a few have specific output to grep for (noted inline).

1. Action (--features=action)

Covers the action runtime end-to-end: pure handlers, async handlers, control-flow nodes, state machines, score-based selectors, and timers.

sh
cargo run --example hello_world     --features=action       # prints "Hello, world!"
cargo run --example simple_action   --features=action       # caller-entity lookup
cargo run --example behavior_tree   --features=action       # sequence + log
cargo run --example state_machine   --features=action       # RunNext jumps
cargo run --example repeat_while    --features=action       # loop + condition
cargo run --example utility_ai      --features=action       # HighestScore
cargo run --example long_running    --features=action       # 1.3s timer chain
cargo run --example malenia         --features=action,rand  # BT + utility AI
2. Scripting (--features=quickjs)
sh
cargo run --example scripting --features=quickjs             # JS Script<I,O>
3. Router (--features=router,markdown / plus extras)

CLI router server, persisted router, and the codegen pipeline.

sh
cargo run --example router           --features=router,markdown
cargo run --example router           --features=router,markdown -- about
cargo run --example cli              --features=router,quickjs -- greet --name=world
# the persisted scene caches the route scripts, so regenerate it after any change
# to script authoring or to a registered type
cargo run --example router_serde     --features=router,quickjs,template_serde -- --new
cargo run --example router_serde     --features=router,quickjs,template_serde
cargo run --example router_serde     --features=router,quickjs,template_serde -- greet --name=world
# rsx_site is a crate, not a root example: generate its routes, then serve. It
# scans typed pages, markdown content and a server action from three collections.
cargo run -p rsx_site --no-default-features --features codegen   # regenerate src/codegen/
cargo run -p rsx_site                                            # http server (default)
cargo run -p rsx_site --features cli -- guide --accept=text/html # render one route to stdout
4. Todo (--features=router,json)

Round-trip the todo document: list → create → list → delete → list.

sh
cargo run --example todo --features=router,json -- list
cargo run --example todo --features=router,json -- create --body='{"description":"smoke test","done":false}'
cargo run --example todo --features=router,json -- list
cargo run --example todo --features=router,json -- delete --body=0
5. Net (--features=net,ureq,native-tls / --features=http_server)

http_client hits example.com and asserts on the response body — skip if offline.

sh
cargo run --example http_client --features=net,ureq,native-tls

For the server side, run in background and probe with curl:

sh
# launch
cargo run --example http_server --features=http_server     # background
curl -s http://localhost:8337                              # expect 200 + body
curl -s http://localhost:8337?name=billy
# kill the background pid

Same pattern for templating (--features=http_server) and style (--features=http_server,style).

6. Per-crate examples

These belong to a specific crate so they need -p.

sh
cargo run -p beet_core --example runner                    # custom test runner
cargo run -p beet_core --example tracing                   # PrettyTracing init
cargo run -p beet_ui   --example render_simple             # oneshot terminal render
cargo run -p beet_ui   --example inline_formatting         # block + inline runs
cargo run -p beet_ui   --example reactive                  # prints "success"
cargo run -p beet_ui   --example build_css   --features=style       # writes target/examples/style/*
cargo run -p beet_ml   --example hello_ml_basic                     # downloads bert (~90MB, slow first run)
cargo run -p beet_ml   --example hello_rl_basic --features=bevy_default
7. Workspace ML (--features=examples,ml)

The examples,ml feature only gates windowed scene code (now scene modules in beet_extra, not runnable --example targets), so there is no self-terminating CLI smoke here. The runtime ML smoke lives in the crate (hello_ml_basic, section 6); this feature's compilation is covered by the skip-set check below (and is the only coverage, since beet_extra is excluded from the test crates).

8. BSX scenes (beet --main=<file>.bsx)

The no-code .bsx scenes run through the installed beet CLI (when editing rust, cargo run -p beet-cli --features=.. -- <args> instead, so the scenes run against the working tree). Each entry documents its own beet --main=.. command in its header, and an entry that declares its hard requirements with <RequireCfg> fails fast on a leaner binary, naming what is missing. The self-terminating ones render and exit:

A documented command never carries --features: that is the entry's own <RequireCfg>'s job. The binary still has to link the capability though, and the demo scenes name actions from beet_extra, which is the extra cargo feature. Build the CLI once with what the set needs and run everything against it:

sh
cargo build -p beet-cli --features=extra        # every scene below except the ml one
cargo build -p beet-cli --features=extra,ml     # adds `hello_ml.bsx`

A binary without extra does not fail fast on these entries the way hello_ml.bsx does: none of them declare <RequireCfg cfg="feature:extra"/>, so instead of a named missing capability you get a spread warning (skipping spread 'SayHello') and then No Action<(), ()>. Worth fixing in the entries; until then, read that pair of messages as "rebuild with extra".

Beware the ml build specifically: it pulls winit and bevy_render, so every scene brings up a wgpu device and compiles compute pipelines whether or not it needs a GPU. On an NVIDIA host that intermittently segfaults inside libnvidia-glcore during create_compute_pipeline, on bevy's async compute thread, which has nothing to do with the scene. Use the extra-only binary for everything but the ml scene.

sh
beet --main=examples/hello                                       # prints "hello world"
beet --main=examples/action/behavior_tree.bsx                    # sequence + log
beet --main=examples/ml/hello_ml.bsx                             # logs "NearestSentence chose: ..."
beet --main=examples/calculator/main.bsx --server=cli add --a=3 --b=4   # result: 7

The rest of examples/action/*.bsx (hello_world, simple_action, long_running, repeat_while, state_machine, utility_ai, scripting, world_script) are also self-terminating and worth a sweep.

Skip: examples/spatial/*.bsx and examples/ml/frozen_lake_*.bsx (windowed), examples/thread/*.bsx (need an LLM key), examples/bsx_site/main.bsx (HTTP server; verify with beet --main=examples/bsx_site --server=cli instead).

Every scene in examples/action/ exits 0. A () -> Outcome load exits zero once it resolves whatever the outcome (an outcome is a branch, not an error), so malenia.bsx and repeat_while.bsx, whose <Repeat> ends by returning its body's fail, report a completed run. Only a scene carrying {OutcomeOverload{error_on_fail:true}} turns a Fail into a nonzero exit.

Show full SKILL.md (301 more words)Show less

Not Verifiable Via CLI (skip)

Documented so future passes don't waste time on them:

  • Spatial / ML scenes: flock, seek, fetch, frozen_lake_run etc. are no longer --example targets — they live as scene modules in beet_extra, reached only by building the examples,spatial / examples,ml features.
  • Thread scenes: chat, multi_agent, oneshot, persistent_chat, tool_call, self_evolving, coding_agent are .bsx markup scenes under examples/thread/, not --example targets. Several also need an LLM key (OPENAI_API_KEY / BEDROCK_*).
  • Interactive TUI/stdin: ui/term_input, ui/tui, ui/state — real examples that block on stdin.
  • Browser required: ui/crud, ui/syntax_highlighting, ui/media_renderer (interactive output).
  • Needs sshd: ssh_server, ssh_client, ssh_tui.

A pure compile check is still useful for the skipped set. The first three cover beet_extra and the feature gates, which no test crate compiles:

sh
cargo check -p beet --features=examples,ml       # beet_extra ML scenes (fetch etc.)
cargo check -p beet --features=examples,spatial  # beet_extra spatial scenes (flock etc.)
cargo check -p beet --features=thread            # thread scene templates
cargo check -p beet-cli --features infra         # examples/infra/*.bsx deploy templates

Instructions

  1. Begin a fresh run by overwriting the scratch file:
    sh
    : > .agents/tmp/scratch.txt
  2. Walk through sections 1–8 and the skip-set compile checks in order, appending each invocation's output:
    sh
    timeout 60 cargo run --example hello_world --features=action 2>&1 | tee -a .agents/tmp/scratch.txt
  3. On a failing example, isolate it with --features=… matching the workspace declaration, fix using a subagent if the fault is non-trivial, then rerun just that example before moving on.
  4. For server examples, launch with run_in_background, probe with curl, kill the background pid before continuing.
  5. After all sections pass, re-grep the scratch file for error, warning, panicked, and FAIL to catch anything missed. Fix any warnings encountered.
  6. Once the full set passes again, provide a comprehensive summary of what changed and which examples were touched.

Success

The smoke set passes when every command in sections 1–8 and the skip-set compile checks exit 0 (or, for the server probes, the curl returns the expected body) and the scratch output contains no error/warning/panicked lines beyond the tracing example's own demo WARN/ERROR. Every other one is fixed, whoever wrote the code it points at.

© mrchantey, Apache-2.0. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

Just SKILL.md in .agents/skills/test-examples of mrchantey/beet.

Open the folder on GitHubat commit 38648b8

Compare with similar skills

Test Examples next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Test Examples compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Test Examples this skillmrchantey/beet135—~2.9kAutomated safety check: PassApache-2.0
TDD WorkflowhellangleZ/burn-in-cceverywhere-ralph11211 repos~2.4kAutomated safety check: PassNone
Testing OpenLogi UIAprilNEA/OpenLogi23k—~1.1kAutomated safety check: PassApache-2.0
Go Testingcxuu/golang-skills1701 repos~1.3kAutomated safety check: PassApache-2.0
Contractssamchon/nestia2.2k—~1.3kAutomated safety check: PassMIT
Cohesion Over TestabilityEpicenterHQ/epicenter4.8k—~2kAutomated safety check: PassCustom licence

Similar skills

  • TDD Workflow

    hellangleZ/burn-in-cceverywhere-ralph

    A skill your agent uses when writing new features, fixing bugs, or refactoring code.

    112 GitHub starsUsed in 11 repos~2.4k tokens
    Testing & QAAuto-check passed
  • Testing OpenLogi UI

    AprilNEA/OpenLogi

    Verifies OpenLogi's native GPUI interface with focused tests, the component gallery and a mock agent, choosing the evidence that fits each change.

    23k GitHub stars~1.1k tokensUpdated 4 days ago
    Testing & QAAuto-check passed
  • Go Testing

    cxuu/golang-skills

    A skill your agent uses when writing, reviewing, or improving Go test code — including table-driven tests, subtests, parallel tests, test helpers, test doubles, and assertions with cmp.Diff.

    170 GitHub starsUsed in 1 repo~1.3k tokens
    Testing & QAAuto-check passed
  • Contracts

    samchon/nestia

    Defines self-acknowledgments for production declarations and tests.

    2.2k GitHub stars~1.3k tokensUpdated yesterday
    Testing & QAAuto-check passed
  • Cohesion Over Testability

    EpicenterHQ/epicenter

    Collapse test-shaped production boundaries while preserving behavior and coverage.

    4.8k GitHub stars~2k tokensUpdated today
    Testing & QAAuto-check passed
  • JS-in-HTML Testing

    liaohch3/claude-tap

    Tests JavaScript embedded in an HTML file in two layers: pytest checks of the logic ported to Python, and Playwright runs in a real browser for the DOM.

    3.3k GitHub stars~924 tokensUpdated 16 days ago
    Testing & QAAuto-check passed

More from mrchantey/beet

All 11 skills in this repo
  • Downstream Create

    mrchantey/beet

    Scaffold a new beet downstream repo (beet<name, a sibling of this checkout depending on the beet facade by path) from the template, register it with downstream-sync and run the first sync.

    135 GitHub stars~709 tokensUpdated yesterday
    Auto-check passed
  • Test The Works

    mrchantey/beet

    Maximal project health pass. An agent skill from mrchantey/beet.

    135 GitHub stars~1k tokensUpdated yesterday
    Auto-check passed
  • Docs Rust Sweep

    mrchantey/beet

    Comprehensive public-api quality pass over the core crates, reducing pub visibility and documenting every public item to a deny(missingdocs) standard.

    135 GitHub stars~590 tokensUpdated yesterday
    Auto-check passed
  • Docs Site

    mrchantey/beet

    Write and edit the public website docs in site/routes/docs (beet.org/docs).

    135 GitHub stars~672 tokensUpdated yesterday
    Auto-check passed
  • Infra Deploy

    mrchantey/beet

    The release process for the beet website. An agent skill from mrchantey/beet.

    135 GitHub stars~16k tokensUpdated yesterday
    Auto-check: warnings
  • Downstream Sync

    mrchantey/beet

    Refresh every downstream repo (beetatproto, beetconnect, beetegress, beetesp, beeteval) with beet's AGENTS.md marker block, its MIT/Apache licenses and shared config files.

    135 GitHub stars~1.1k tokensUpdated yesterday
    Auto-check passed

Categories

Questions about Test Examples

What does Test Examples do?

Run the workspace example smoke set to catch regressions unit tests miss, covering every feature gate combination with self-terminating examples and curl-probed servers. Test Examples is an agent skill from mrchantey/beet. Run the workspace example smoke set to catch regressions unit tests miss, covering every feature gate combination with self-terminating examples and curl-probed servers.

When should I use Test Examples?

Test Examples fits situations like: verify the examples; tasks that involve Unit testing.

How do I install Test Examples in Claude Code?

Run `npx skills add mrchantey/beet --skill test-examples -a claude-code`. Or copy the skill folder (.agents/skills/test-examples in mrchantey/beet) into .claude/skills/test-examples in your project. Claude Code loads it when a task matches its description.

How do I install Test Examples in Codex?

Run `npx skills add mrchantey/beet --skill test-examples -a codex`. Or copy the skill folder (.agents/skills/test-examples in mrchantey/beet) into .agents/skills/test-examples in your project. Codex loads it when a task matches its description.

Can I use Test Examples in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add mrchantey/beet --skill test-examples -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/test-examples, .gemini/skills/test-examples, .github/skills/test-examples and .opencode/skills/test-examples in your project.

What does Test Examples need to run?

Going by SKILL.md and its folder, Test Examples needs the command-line tools its instructions call (cargo and curl) and credentials named OPENAI_API_KEY. Our summary lists: A credential in OPENAI_API_KEY.

Does Test Examples access the network?

SKILL.md contains no URLs. Its commands use curl, which can reach the network depending on how they are called. This is read from the text; nothing was executed.

Is Test Examples safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Test Examples use?

Test Examples is published under the Apache-2.0 licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Test Examples use?

About 2.9k tokens (SKILL.md is roughly 12k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Test Examples?

Skills that share tags, products or a category with Test Examples: TDD Workflow (hellangleZ/burn-in-cceverywhere-ralph, 112 stars), Testing OpenLogi UI (AprilNEA/OpenLogi, 23k stars), Go Testing (cxuu/golang-skills, 170 stars) and Contracts (samchon/nestia, 2.2k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Test Examples?

mrchantey (a GitHub user) maintains it in mrchantey/beet, which has 135 GitHub stars. The repository holds 11 skills in this directory. The repository was last updated on October 7, 2026.

Source: mrchantey/beet on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.