Agent skill

Roam Measurement Design

by Cranot in Cranot/roam-code

Design or interpret a comparison of Roam performance, detector accuracy, retrieval or workflow value.

Apache-2.0Auto-check passedDevelopment

Install Roam Measurement Design

skills CLI
$ npx skills add Cranot/roam-code --skill roam-measurement-design -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install Cranot/roam-code roam-measurement-design --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/Cranot/roam-code.git skills-src && mkdir -p .claude/skills && cp -r skills-src/.agents/skills/roam-measurement-design .claude/skills/roam-measurement-design && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
roam-measurement-design
GitHub stars
517
Token cost
~1.2k tokens
SKILL.md length
589 words
Files
2
Skills in repo
9
Repo updated
First seen
Licence
Apache-2.0

At a glance

Design or interpret a comparison of Roam performance, detector accuracy, retrieval or workflow value.

  • Choose a fair measurement and bound its resulting claim
  • SKILL.md covers Choose the question before the…, Match the observation to the… and Produce a bounded decision…
  • Instructions only: no scripts, shell commands, URLs or credentials in SKILL.md
  • Not for ordinary regression tests

What it does

Roam Measurement Design is an agent skill from Cranot/roam-code. Design or interpret a comparison of Roam performance, detector accuracy, retrieval or workflow value. Use to choose a fair measurement and bound its resulting claim, not for ordinary regression tests, product copy, release approval or a new benchmark platform.

Its SKILL.md is about 1.2k tokens, which your agent loads only when the skill is triggered. The skill folder holds 2 other files (for example `agents/openai.yaml`).

It sits in Development, covering Copywriting. It works with Model Context Protocol. The repository describes itself as: Local codebase intelligence CLI + MCP server for AI coding agents: SQLite code graph, 28 languages, 287 commands, 246 MCP tools, change-safety gates, audit evidence, zero API keys. The licence is Apache-2.0.

When your agent uses it

  • Choose a fair measurement and bound its resulting claim
  • Not for ordinary regression tests
  • Release approval
  • A new benchmark platform

Example prompts

  • “/roam-measurement-design”

What it can do on your machine

Read from SKILL.md and the folder at commit f0bdb63. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    No scripts in the folder and no shell commands in SKILL.md.

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Roam Measurement Design loads about 1.2k tokens when it runs. Until then it costs about 71 tokens; SKILL.md has 589 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~71
When it runs · the whole SKILL.md, loaded when a task matches
~1.2k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from Cranot/roam-code at commit f0bdb63, republished under its Apache-2.0 licence (© Cranot). 589 words, ~1,206 tokens.

Download SKILL.mdSave it as .claude/skills/roam-measurement-design/SKILL.md (or your agent's skills folder). This skill also uses 1 other file; get the full folder from GitHub.
name
roam-measurement-design
description
Design or interpret a comparison of Roam performance, detector accuracy, retrieval or workflow value. Use to choose a fair measurement and bound its resulting claim, not for ordinary regression tests, product copy, release approval or a new benchmark platform.

Roam measurement design

Turn a proposed improvement into a comparison that can change a decision. An exploratory result can justify the next check without establishing general benefit. A good design can be retained; more trials, labels or machinery are not automatically better evidence.

Choose the question before the machinery

Use the maintained method, not a copied checklist. Read the sections applicable to the selected claim:

  • Comparisons: Comparing analytical approaches in docs/concepts/verification-evidence.md owns controls and acceptance rules.
  • Harness/model runs: that guide's Benchmark accounting owns run populations.
  • Detector/graph accuracy: docs/concepts/detector-evidence.md and Limits an agent must respect in docs/understanding-roam.md own applicability and the requirements for independent/generalized claims, beyond local exploration.
  • Performance/workflow value: Evidence and acceptance in internal/WORK-GUIDE.md; use the bets/accounting in internal/ROAM-NORTH-STAR.md when choosing a value experiment. Public numbers also need the orientation's Numbers requirements.

Read known internal/ paths directly before declaring them unavailable; default file listings can omit these gitignored authorities. Missing private strategy is a named limit, not a veto of an adequate bounded technical comparison. Frozen sources may stand in for repository reads in a trial.

Name the decision, intended population and binding resource. Distinguish an exploratory local comparison, a confirmatory study and a retrospective analysis. Fix prospective acceptance and stop/narrowing criteria before assignment; describe retrospective choices honestly instead of pretending they were preregistered. Use an owner-agreed material margin when needed, not an invented percentage, sample size or composite score. Study design is not approval to run paid trials, contact users, access other repos or change production.

Compare the proposed intervention with a capable incumbent and a simpler alternative, using the comparison guide's controls; qualify deliberate changes separately. Separate a warm-query microbenchmark from indexing, startup, integration, fallback and maintenance costs at the repetition rate that matters. Verify output and accepted-outcome equivalence; removing work can be faster without being an improvement. Internal code repair needs no customer or efficacy study.

Show full SKILL.md (280 more words)Show less

Match the observation to the claim

Match the population to the question: emitted findings for finding precision, independently sampled eligible sites for bounded recall, accepted task outcomes for workflow value. Apply the source's sampling, adjudication and applicability requirements; developer-selected examples remain exploratory, not general accuracy.

Keep failed attempts, false refusals and missing observations visible under the applicable accounting method, with each metric's known denominator. A zero is not a missing value. Separate resources instead of converting tokens, seconds and owner attention into an unexplained score.

Apply the comparison guide's development/confirmation separation and healthy-case controls. Choose the smallest study that can resolve the stated decision; preserve the cheaper method when added complexity does not earn its cost.

Produce a bounded decision record

Before running, record the bounded protocol and unresolved choices needed to execute; reuse an adequate design. Afterwards preserve exact artifacts, identities, dates, denominators and metric definitions. State uncertainty appropriate to the design: neither a small identical success count nor a best cell proves parity or general benefit. Do not invent intervals from insufficient records.

Write what the result establishes, what it leaves unknown, and whether to retain, narrow, investigate or stop. A descriptive internal result can be useful without a publishable efficacy claim. A new public numerical claim also needs the project's reproducibility/publication requirements; preserve historical records rather than silently updating their numbers. Product wording and release approval are separate jobs, not consequences of a favorable comparison.

Use existing measurement/accounting tools after verifying their actual inputs and coverage. Do not build a runner or run an expensive study merely to finish a design request. Keep protocols and dated observations private under internal/ unless public documentation is explicitly in scope.

© Cranot, Apache-2.0. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

SKILL.md and 1 other file in .agents/skills/roam-measurement-design of Cranot/roam-code.

  • SKILL.md
  • agents/openai.yaml

Open the folder on GitHubat commit f0bdb63

Compare with similar skills

Roam Measurement Design next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Roam Measurement Design compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Roam Measurement Design this skillCranot/roam-code517—~1.2kAutomated safety check: PassApache-2.0
Writeohad6k/emulo2941 repos~488Automated safety check: PassMIT
Score Harnessruvnet/metaharness696—~613Automated safety check: PassMIT
Testing With API Mocksstacklok/toolhive-studio171—~1.5kAutomated safety check: PassApache-2.0
Plan A Feature To Confluencetestdouble/han281—~4.2kAutomated safety check: PassMIT
Chrome Devtools MCPmanagedcode/dotnet-skills486—~2.2kAutomated safety check: PassMIT

Similar skills

  • Write

    ohad6k/emulo

    A skill your agent uses for marketing, social, replies, product copy, launch copy, and writing in the user's voice when their Emulo writing profile should guide the task.

    294 GitHub starsUsed in 1 repo~488 tokens
    Agent WorkflowsAuto-check passed
  • Score Harness

    ruvnet/metaharness

    5-dimension scorecard (0-100, grade A/B/C/F) for a scaffolded harness.

    696 GitHub stars~613 tokensUpdated yesterday
    DevelopmentAuto-check passed
  • Testing With API Mocks

    stacklok/toolhive-studio

    Start here for all API mocking in tests. An agent skill from stacklok/toolhive-studio.

    171 GitHub stars~1.5k tokensUpdated yesterday
    DevelopmentAuto-check passed
  • Builds a feature specification from scratch with plan-a-feature and publishes it to a user-specified Confluence location, posting the spec as a parent page and each companion artifact (decision log…

    281 GitHub stars~4.2k tokensUpdated 10 days ago
    DevelopmentAuto-check passed
  • Chrome Devtools MCP

    managedcode/dotnet-skills

    Use Chrome DevTools MCP from .NET agents and .NET-focused repos to inspect, debug, and automate Chrome through an MCP client.

    486 GitHub stars~2.2k tokensUpdated yesterday
    DevelopmentAuto-check passed
  • Analyze Logs

    activepieces/activepieces

    Analyze application logs from the .evlog/logs/ directory. An agent skill from activepieces/activepieces.

    25k GitHub starsUsed in 1 repo~1.6k tokens
    DevelopmentAuto-check passed

More from Cranot/roam-code

All 9 skills in this repo
  • Roam

    Cranot/roam-code

    Codebase comprehension via roam-code CLI. An agent skill from Cranot/roam-code.

    517 GitHub stars~2.4k tokensUpdated yesterday
    Auto-check passed
  • Roam Agent Guidance

    Cranot/roam-code

    Review and maintain the technical instructions Roam gives coding agents: shipped skills, tool descriptions, preset guidance and generated instruction blocks.

    517 GitHub stars~1.3k tokensUpdated yesterday
    Auto-check passed
  • Roam Evidence Hardening

    Cranot/roam-code

    Investigate and harden Roam detectors, CLI/MCP result contracts, and evidence consumers when dogfooding or correcting incomplete, misleading, or inconsistent analysis.

    517 GitHub stars~1.5k tokensUpdated yesterday
    Auto-check passed
  • Roam Lesson Maintenance

    Cranot/roam-code

    Turn a reproduced or recurring Roam failure into a durable correction, regression control or narrowly scoped project skill, and reconcile conflicting lessons.

    517 GitHub stars~1.2k tokensUpdated yesterday
    Auto-check passed
  • Roam Milestone Planning

    Cranot/roam-code

    Define or revise a Roam product milestone and its engineering, adoption and offer-readiness sequence.

    517 GitHub stars~1.2k tokensUpdated yesterday
    Auto-check passed
  • Roam Release Readiness

    Cranot/roam-code

    Qualify a Roam commit or accumulated batch for package, website, or server publication; reconcile source, tests, review, CI and deployed identities, and identify the next authorized release step.

    517 GitHub stars~1.7k tokensUpdated yesterday
    Auto-check passed

Questions about Roam Measurement Design

What does Roam Measurement Design do?

Design or interpret a comparison of Roam performance, detector accuracy, retrieval or workflow value. Roam Measurement Design is an agent skill from Cranot/roam-code. Design or interpret a comparison of Roam performance, detector accuracy, retrieval or workflow value.

When should I use Roam Measurement Design?

Roam Measurement Design fits situations like: choose a fair measurement and bound its resulting claim; not for ordinary regression tests; release approval; A new benchmark platform.

How do I install Roam Measurement Design in Claude Code?

Run `npx skills add Cranot/roam-code --skill roam-measurement-design -a claude-code`. Or copy the skill folder (.agents/skills/roam-measurement-design in Cranot/roam-code) into .claude/skills/roam-measurement-design in your project. Claude Code loads it when a task matches its description.

How do I install Roam Measurement Design in Codex?

Run `npx skills add Cranot/roam-code --skill roam-measurement-design -a codex`. Or copy the skill folder (.agents/skills/roam-measurement-design in Cranot/roam-code) into .agents/skills/roam-measurement-design in your project. Codex loads it when a task matches its description.

Can I use Roam Measurement Design in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add Cranot/roam-code --skill roam-measurement-design -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/roam-measurement-design, .gemini/skills/roam-measurement-design, .github/skills/roam-measurement-design and .opencode/skills/roam-measurement-design in your project.

What does Roam Measurement Design need to run?

SKILL.md names no scripts, command-line tools or credentials: Roam Measurement Design is instructions for the agent only.

Does Roam Measurement Design access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is Roam Measurement Design safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Roam Measurement Design use?

Roam Measurement Design is published under the Apache-2.0 licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Roam Measurement Design use?

About 1.2k tokens (SKILL.md is roughly 4.8k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Roam Measurement Design?

Skills that share tags, products or a category with Roam Measurement Design: Write (ohad6k/emulo, 294 stars), Score Harness (ruvnet/metaharness, 696 stars), Testing With API Mocks (stacklok/toolhive-studio, 171 stars) and Plan A Feature To Confluence (testdouble/han, 281 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Roam Measurement Design?

Cranot (a GitHub user) maintains it in Cranot/roam-code, which has 517 GitHub stars. The repository holds 9 skills in this directory. The repository was last updated on October 10, 2026.

Source: Cranot/roam-code on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.