Official agent skill

Trailblaze Validate Oob

by block in block/trailblaze

A skill your agent uses when validating or evaluating Trailblaze's out-of-box (OOB) user experience — does what the trailblaze skill claims actually match the real installed CLI?

OfficialApache-2.0Auto-check passedFrontend & Design

Install Trailblaze Validate Oob

skills CLI
$ npx skills add block/trailblaze --skill trailblaze-validate-oob -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install block/trailblaze trailblaze-validate-oob --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/block/trailblaze.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/trailblaze-validate-oob .claude/skills/trailblaze-validate-oob && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
trailblaze-validate-oob
GitHub stars
321
Token cost
~2.3k tokens
SKILL.md length
1,276 words
Files
1
Skills in repo
4
Repo updated
First seen
Licence
Apache-2.0

At a glance

A skill your agent uses when validating or evaluating Trailblaze's out-of-box (OOB) user experience — does what the trailblaze skill claims actually match the real installed CLI?

  • Works in 6 steps: Install the binary fresh → Spawn a fresh-context subagent → Give the subagent a concrete scenario → …
  • Evaluating Trailblazes out-of-box (OOB) user experience — does what the trailblaze skill claims actually match the real installed CLI?
  • SKILL.md covers When to run a validation pass, What "good" looks like, Methodology and What to NOT do, plus 2 more sections
  • Instructions only: no scripts, shell commands, URLs or credentials in SKILL.md

What it does

Trailblaze Validate Oob is an agent skill from block/trailblaze, published by the product's own GitHub organization. Use when validating or evaluating Trailblaze's out-of-box (OOB) user experience — does what the trailblaze skill claims actually match the real installed CLI? Triggers on requests to "validate the Trailblaze skill", "test the Trailblaze OOB experience", "evaluate Trailblaze UX", "check if the skill matches the CLI", "audit the Trailblaze CLI for new-user friction", or running an OOB regression check after the framework changes.

Its SKILL.md is about 2.3k tokens, which your agent loads only when the skill is triggered. It is a single SKILL.md file with no bundled scripts.

It sits in Frontend & Design, covering UX design. The repository describes itself as: 🥾 AI-Driven UI Testing Framework with Recorded Trails. The licence is Apache-2.0.

When your agent uses it

  • Evaluating Trailblazes out-of-box (OOB) user experience — does what the trailblaze skill claims actually match the real installed CLI?
  • Requests to validate the Trailblaze skill
  • Test the Trailblaze OOB experience
  • Evaluate Trailblaze UX

Example prompts

  • “validate the Trailblaze skill”
  • “test the Trailblaze OOB experience”
  • “evaluate Trailblaze UX”
  • “/trailblaze-validate-oob”

Workflow steps

6 steps, taken from the step headings in SKILL.md.

  1. Install the binary fresh
  2. Spawn a fresh-context subagent
  3. Give the subagent a concrete scenario
  4. Subagent prompt template
  5. Triage findings
  6. Run a second pass after fixes

What it can do on your machine

Read from SKILL.md and the folder at commit ec49f40. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    No scripts in the folder and no shell commands in SKILL.md (its code samples are bash).

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Trailblaze Validate Oob loads about 2.3k tokens when it runs. Until then it costs about 114 tokens; SKILL.md has 1,276 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~114
When it runs · the whole SKILL.md, loaded when a task matches
~2.3k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from block/trailblaze at commit ec49f40, republished under its Apache-2.0 licence (© block). 1,276 words, ~2,276 tokens.

Download SKILL.mdSave it as .claude/skills/trailblaze-validate-oob/SKILL.md (or your agent's skills folder).
name
trailblaze-validate-oob
description
Use when validating or evaluating Trailblaze's out-of-box (OOB) user experience — does what the `trailblaze` skill claims actually match the real installed CLI? Triggers on requests to "validate the Trailblaze skill", "test the Trailblaze OOB experience", "evaluate Trailblaze UX", "check if the skill matches the CLI", "audit the Trailblaze CLI for new-user friction", or running an OOB regression check after the framework changes.

Validating Trailblaze's out-of-box experience

This skill encodes the methodology for honestly evaluating whether Trailblaze's user-facing CLI matches what a fresh agent / new user would intuitively expect. It's the framework's own UX regression test, run by an agent simulating cold-start use.

The goal is to catch counter-intuitive defaults, broken claims in the trailblaze skill, and friction that the team has gone CLI-blind on.

When to run a validation pass

  • After any change to the trailblaze CLI surface (commands, flags, default behavior)
  • After any change to the trailblaze Claude skill (the sibling trailblaze/SKILL.md) or its references
  • Periodically (monthly?) even with no changes, to catch drift
  • When a new rung of the adoption ladder gets added to the skill

What "good" looks like

After a validation pass, a fresh agent following only the trailblaze skill should be able to complete the scenario with zero surprises. Specifically:

  • Every command the skill says to run actually exists with the documented flags and argument shape
  • Every output the skill describes matches what the CLI actually prints
  • The basic loop (drive → save → replay → inspect) works end-to-end with no undocumented gotchas
  • Failure modes the skill mentions are real failure modes; the recovery steps actually recover
  • Errors the user might hit have clear hint: lines pointing at the fix
  • No "workaround" paragraphs in the skill that exist to paper over CLI behavior the user wouldn't naturally guess

If the validation surfaces something the agent couldn't have guessed from the skill alone, that's either a skill update OR a CLI fix — and the latter is usually preferable.

Methodology

Step 1: Install the binary fresh

Use the actual end-user install path, NOT ./trailblaze from the repo (which goes through Gradle on every call and isn't what users hit):

bash
./scripts/install-trailblaze-source.sh

The script builds the uber JAR and installs it under ~/.trailblaze/install/ with a symlink onto the system PATH. The specific symlink location depends on your platform (homebrew on macOS, /usr/local/bin/ on Intel macs / Linux, …) — defer to the script's stdout for the exact path rather than memorizing one. Then verify:

bash
which trailblaze         # confirms which binary the shell will pick up
trailblaze --version     # records the version under test (capture this in the report)

If the script fails (missing JDK, Gradle error, symlink-permission failure), that's itself a finding: file it as an "OOB blocker" and spawn a CLI fix chip before trying to run the validation — don't shim around it from inside the skill.

Step 2: Spawn a fresh-context subagent

The subagent must have no prior Trailblaze knowledge. Their only source of truth is the trailblaze skill file. They must not infer commands from the repo's source code or from any other docs.

The right tool for this is a general-purpose agent (read-only is sufficient since the validation only runs --help and discovery commands, never destructive actions).

If the subagent can't proceed at this step — no device connected, target app not installed, network or auth required, etc. — that's a gap too: an OOB experience that requires undocumented setup is itself a finding. Record what setup was needed and what error the subagent hit, file it the same way as an OOB blocker, and resume the validation once the setup gap is closed.

Step 3: Give the subagent a concrete scenario

A scenario must:

  • Be a realistic user task ("drive an Android emulator", "save a flow as a trail", "replay a trail and inspect what happened")
  • Be doable using ONLY commands the skill describes
  • Have a clear "I would now act on a device" stopping point — the validation is read-only; don't actually drive devices or modify state

Scenarios that have worked well:

  • Rung 1: "Drive a device. List devices, snapshot one, look at the toolbox to find available actions, identify how you'd act on an element. Stop before actually executing the action."
  • Rung 2: "Save the current session as a .trail.yaml. Then replay one. Then generate a report. Then look up past results."
  • Rung 3: (when written) "Compose your own agent surface — write a custom typed tool, install a trailmap, curate what tools the agent sees."
Step 4: Subagent prompt template

The validation prompt should always:

  1. Tell the subagent it's a Claude Code session that just loaded the skill — no prior Trailblaze knowledge
  2. Specify the installed trailblaze binary on PATH as the test surface (NOT ./trailblaze)
  3. Give the concrete scenario
  4. Instruct them to follow ONLY what the skill says; gaps are gaps
  5. Require structured output: Inaccuracies, Gaps, Ambiguities, What works, Suggestions
  6. Cite quoted skill text and actual CLI output side-by-side for each finding
Show full SKILL.md (546 more words)Show less
Step 5: Triage findings

For each finding the subagent reports:

Finding typeRight action
Skill claims X; CLI does YUpdate the skill OR spawn a CLI fix chip. Lean toward CLI fix if X is the more intuitive design.
Skill is silent; agent had to guessAdd to the skill. Note explicitly if it's a workaround vs intentional behavior.
Skill is ambiguous; agent picked one interpretationTighten the skill text.
Skill claim worked exactly as describedKeep — note what worked so it doesn't accidentally regress.

Bias toward fixing the CLI, not the skill. If a fresh agent guessed differently than the current CLI, the agent is usually modeling the intuitive design — and the CLI is the thing diverging from intuition.

Step 6: Run a second pass after fixes

After applying fixes (either in the skill or via spawned CLI chips), run the validation again with a different fresh-context subagent. Two passes catch issues the first pass's fixes introduced and confirm the fixes hold against another cold reader.

Don't reuse the same subagent — context contamination defeats the "fresh new user" model.

What to NOT do

  • Don't run validation against ./trailblaze — that's the Gradle-wrapped dev launcher, not the end-user binary. Findings there ("first call triggers a Gradle build", "needs local.properties") are dev-environment concerns, not OOB UX problems.
  • Don't pre-bias the subagent with answers. Don't say "verify that tap is the tool name" — let them discover it from the skill and the CLI.
  • Don't have the subagent execute destructive actions. No driving devices, no modifying state, no committing files. Read-only validation only — --help, device list, snapshot, etc.
  • Don't ignore a subagent finding because "the current CLI design is what it is". That's how counter-intuitive designs calcify.

What a healthy validation report looks like

A useful report has all of:

  • Inaccuracies with quoted skill text + actual CLI output for each
  • Verified claims — explicitly list what held up, so the team knows what's safe to keep
  • Gaps the subagent had to guess past
  • Suggestions — concrete edits to skill OR concrete CLI fixes
  • Cap around 500 words; brevity forces precision

If the report is just "looks fine to me," the methodology wasn't adversarial enough — re-prompt with a more pointed scenario.

After validation: distribution

Findings flow three ways — the first one is non-optional, the other two depend on what surfaced:

  1. Durable record of the pass itself. Before spawning any chips or editing any skill, post a short summary to a persistent surface — a GitHub issue on the project's issue tracker (preferred), or the relevant team chat thread, or a tracked markdown file. Include: date, the rung exercised, the trailblaze --version you captured in Step 1, a one-line outcome ("clean", "N findings, M chipped"), and links to any chips or follow-up PRs. The chat transcript of the validation is not a durable artifact — without this, "did we last validate after the X CLI change?" is unanswerable.
  2. Skill updates — edit the sibling trailblaze/SKILL.md directly, run the sensitive-terms scanner, commit, PR.
  3. CLI fix chips — spawn a focused chip for the framework behavior change, with the validation finding as motivation in the chip's prompt.

Skill updates and CLI chips should reference the validation pass that surfaced the finding (the issue/Slack URL from #1), so future contributors can trace why a change was made.

© block, Apache-2.0. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

Just SKILL.md in skills/trailblaze-validate-oob of block/trailblaze.

Open the folder on GitHubat commit ec49f40

Compare with similar skills

Trailblaze Validate Oob next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Trailblaze Validate Oob compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Trailblaze Validate Oob this skillblock/trailblaze321—~2.3kAutomated safety check: PassApache-2.0
Impeccablebestofjs/bestofjs3.1k27 repos~2.6kAutomated safety check: PassMIT
Interface Design for Dashboards and Appsholaboss-ai/holaOS11k3 repos~6kAutomated safety check: PassMIT
Animategrowupanand/ConvoForm1016 repos~1.9kAutomated safety check: PassApache-2.0
Migrate Content Iadocker/docs4.7k—~5.1kAutomated safety check: PassApache-2.0
UX WalkthroughXiaoMi/hiui877—~1.3kAutomated safety check: PassMIT

Similar skills

  • Impeccable

    bestofjs/bestofjs

    A skill your agent uses when the user wants to design, redesign, shape, critique, audit, polish, clarify, distill, harden, optimize, adapt, animate, colorize, extract, or otherwise improve a…

    3.1k GitHub starsUsed in 27 repos~2.6k tokens
    Frontend & DesignAuto-check passed
  • Pushes an agent past generic defaults when designing dashboards, admin panels, SaaS apps and tools, with attention to structure, type, navigation and how data is shown.

    11k GitHub starsUsed in 3 repos~6k tokens
    Frontend & DesignAuto-check passed
  • Animate

    growupanand/ConvoForm

    Review a feature and enhance it with purposeful animations, micro-interactions, and motion effects that improve usability and delight.

    101 GitHub starsUsed in 6 repos~1.9k tokens
    Frontend & DesignAuto-check passed
  • Official

    Handle Hugo docs information-architecture moves: discover old vs new URLs, add front matter aliases (Phase 1), update in-repo links (Phase 2), interactive List 2 resolution and fragment validation…

    4.7k GitHub stars~5.1k tokensUpdated yesterday
    Frontend & DesignAuto-check passed
  • UX Walkthrough

    XiaoMi/hiui

    体验走查 skill。适用于代码库、URL、截图三种输入,输出结构化体验问题报告,并同步生成本地 docx 报告。触发词:体验走查、UX review、交互走查、界面审查、体验问题。

    877 GitHub stars~1.3k tokensUpdated 2 mo ago
    Frontend & DesignAuto-check passed
  • Color Audit

    rome-os/rome

    Audit a design system's color palette against measurable color-science disciplines — WCAG/APCA contrast of declared token pairs, perceptual (OKLCH) ramp uniformity, color-blindness safety of…

    725 GitHub stars~2.7k tokensUpdated today
    Frontend & DesignAuto-check passed

More from block/trailblaze

  • Devlog

    block/trailblaze

    Official

    Write or update a devlog entry in the devlog directory. An agent skill from block/trailblaze.

    321 GitHub stars~586 tokensUpdated yesterday
    Auto-check passed
  • Trailblaze Author

    block/trailblaze

    Official

    A skill your agent uses when turning a captured human demonstration (a Trailblaze App demonstration bundle: demo.yaml + actions.ndjson + per-action screenshots and view hierarchies) into a durable…

    321 GitHub stars~2.8k tokensUpdated yesterday
    Auto-check passed
  • Trailblaze

    block/trailblaze

    Official

    A skill your agent uses when working with Trailblaze — natural-language device control for coding agents across iOS, Android, and web, with replayable .trail.yaml files as the artifact.

    321 GitHub stars~3.4k tokensUpdated yesterday
    Auto-check passed

Questions about Trailblaze Validate Oob

What does Trailblaze Validate Oob do?

A skill your agent uses when validating or evaluating Trailblaze's out-of-box (OOB) user experience — does what the trailblaze skill claims actually match the real installed CLI? Trailblaze Validate Oob is an agent skill from block/trailblaze, published by the product's own GitHub organization. Use when validating or evaluating Trailblaze's out-of-box (OOB) user experience — does what the trailblaze skill claims actually match the real installed CLI?

When should I use Trailblaze Validate Oob?

Trailblaze Validate Oob fits situations like: evaluating Trailblazes out-of-box (OOB) user experience — does what the trailblaze skill claims actually match the real installed CLI?; requests to validate the Trailblaze skill; test the Trailblaze OOB experience; evaluate Trailblaze UX.

How do I install Trailblaze Validate Oob in Claude Code?

Run `npx skills add block/trailblaze --skill trailblaze-validate-oob -a claude-code`. Or copy the skill folder (skills/trailblaze-validate-oob in block/trailblaze) into .claude/skills/trailblaze-validate-oob in your project. Claude Code loads it when a task matches its description.

How do I install Trailblaze Validate Oob in Codex?

Run `npx skills add block/trailblaze --skill trailblaze-validate-oob -a codex`. Or copy the skill folder (skills/trailblaze-validate-oob in block/trailblaze) into .agents/skills/trailblaze-validate-oob in your project. Codex loads it when a task matches its description.

Can I use Trailblaze Validate Oob in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add block/trailblaze --skill trailblaze-validate-oob -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/trailblaze-validate-oob, .gemini/skills/trailblaze-validate-oob, .github/skills/trailblaze-validate-oob and .opencode/skills/trailblaze-validate-oob in your project.

What does Trailblaze Validate Oob need to run?

SKILL.md names no scripts, command-line tools or credentials: Trailblaze Validate Oob is instructions for the agent only.

Does Trailblaze Validate Oob access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is Trailblaze Validate Oob safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Trailblaze Validate Oob use?

Trailblaze Validate Oob is published under the Apache-2.0 licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Trailblaze Validate Oob use?

About 2.3k tokens (SKILL.md is roughly 9.1k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Trailblaze Validate Oob?

Skills that share tags, products or a category with Trailblaze Validate Oob: Impeccable (bestofjs/bestofjs, 3.1k stars), Interface Design for Dashboards and Apps (holaboss-ai/holaOS, 11k stars), Animate (growupanand/ConvoForm, 101 stars) and Migrate Content Ia (docker/docs, 4.7k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Trailblaze Validate Oob?

block (a GitHub organization, an official publisher) maintains it in block/trailblaze, which has 321 GitHub stars. The repository holds 4 skills in this directory. The repository was last updated on October 7, 2026.

Source: block/trailblaze on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.