Agent skill

Cost Trend

by ruvnet in ruvnet/ruflo

Read every docs/benchmarks/runs/.json and surface drift in win rate, latency, escalation rate, and LLM-baseline cost over time

MITAuto-check: notes

Install Cost Trend

skills CLI
$ npx skills add ruvnet/ruflo --skill cost-trend -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install ruvnet/ruflo cost-trend --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/ruvnet/ruflo.git skills-src && mkdir -p .claude/skills && cp -r skills-src/plugins/ruflo-cost-tracker/skills/cost-trend .claude/skills/cost-trend && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
cost-trend
GitHub stars
74k
Token cost
~490 tokens
SKILL.md length
238 words
Files
1
Skills in repo
264
Repo updated
First seen
Licence
MIT

At a glance

Read every docs/benchmarks/runs/.json and surface drift in win rate, latency, escalation rate, and LLM-baseline cost over time

  • Works in 4 steps: Run the trend script from the project root → Inspect the drift summary — first vs… → Inspect the per-run series — one row per… → …
  • SKILL.md covers When to use, Steps and Cross-references
  • Calls node

What it does

Cost Trend is an agent skill from ruvnet/ruflo. Read every docs/benchmarks/runs/.json and surface drift in win rate, latency, escalation rate, and LLM-baseline cost over time

Its SKILL.md is about 490 tokens, which your agent loads only when the skill is triggered. It is a single SKILL.md file with no bundled scripts.

The repository describes itself as: 🌊 The original agent harness. Deploy intelligent multi-player swarms, coordinate autonomous workflows, and build conversational AI systems. Features adaptive memory…. The licence is MIT.

Example prompts

  • “/cost-trend”

Requirements

  • Pre-approved tools (allowed-tools): Bash

Workflow steps

4 steps, taken from the first numbered list in SKILL.md.

  1. Run the trend script from the project root
  2. Inspect the drift summary — first vs last on win rate, avg latency, p99, escalation rate, speedup vs Gemini.
  3. Inspect the per-run series — one row per run, including Sonnet 4.6 + Opus 4.7 baseline latencies if those were enabled (BENCH_ANTHROPIC=1…
  4. Regression flags — the script emits > ⚠ Regression callouts when

What it can do on your machine

Read from SKILL.md and the folder at commit 58e0ae7. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves these tools, so the agent can use them without asking each time:

    • Bash

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Shell commands in SKILL.md call:

    • node

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Cost Trend loads about 490 tokens when it runs. Until then it costs about 35 tokens; SKILL.md has 238 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~35
When it runs · the whole SKILL.md, loaded when a task matches
~490

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check: notes

The automated check noted patterns worth knowing about, such as sudo or a known installer.

  • NotePre-approves every shell command (allowed-tools: Bash)SKILL.md
    allowed-tools: Bash

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from ruvnet/ruflo at commit 58e0ae7, republished under its MIT licence (© ruvnet). 238 words, ~490 tokens.

Download SKILL.mdSave it as .claude/skills/cost-trend/SKILL.md (or your agent's skills folder).
name
cost-trend
description
Read every docs/benchmarks/runs/*.json and surface drift in win rate, latency, escalation rate, and LLM-baseline cost over time
allowed-tools
Bash

Cost Trend

The smoke gate is binary (winRate ≥ 0.80 → pass/fail). The corpus benchmarks captured over time form a curve — and curves catch regressions the gate misses (win rate slowly creeping from 100% to 85% is "still passing" by smoke but a real degradation).

This skill reads every persisted run in docs/benchmarks/runs/*.json and reports first→last deltas plus a per-run series, flagging regressions in win rate or latency.

When to use

  • Before a release — check that the speedup hasn't drifted.
  • After expanding the corpus — verify older runs still hit the same win rate on the new corpus they reflected.
  • After upgrading agent-booster — surface latency / strategy changes.

Steps

  1. Run the trend script from the project root:

    bash
    node plugins/ruflo-cost-tracker/scripts/trend.mjs

    Optional env:

    • TREND_FORMAT=json — emit JSON instead of markdown
    • TREND_LIMIT=10 — consider only the most recent N runs
  2. Inspect the drift summary — first vs last on win rate, avg latency, p99, escalation rate, speedup vs Gemini.

  3. Inspect the per-run series — one row per run, including Sonnet 4.6 + Opus 4.7 baseline latencies if those were enabled (BENCH_ANTHROPIC=1 at run time).

  4. Regression flags — the script emits > ⚠ Regression callouts when:

    • Win rate dropped between first and last run
    • Avg latency rose ≥ 1.5× from first run

Cross-references

  • cost-benchmark — the producer of the run JSONs this skill consumes
  • bench/booster-corpus.json — the corpus version is recorded in each run, so trends across corpus versions remain interpretable
  • docs/benchmarks/runs/latest.json — the most-recent run; smoke step 23 gates on winRate ≥ 0.80 from this file

© ruvnet, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

Just SKILL.md in plugins/ruflo-cost-tracker/skills/cost-trend of ruvnet/ruflo.

Open the folder on GitHubat commit 58e0ae7

Compare with similar skills

Cost Trend next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Cost Trend compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Cost Trend this skillruvnet/ruflo74k—~490Automated safety check: NotesMIT
Benchmarkaffaan-m/ECC276k3 repos~654Automated safety check: PassMIT
Benchmarkaffaan-m/ECC276k—~412Automated safety check: PassMIT
Benchmarkaffaan-m/ECC276k—~330Automated safety check: PassMIT
Cost Trackingaffaan-m/ECC276k1 repos~1.3kAutomated safety check: PassMIT
Gstack Performance Benchmarkgarrytan/gstack136k—~7.3kAutomated safety check: NotesMIT

Similar skills

  • Benchmark

    affaan-m/ECC

    Measure performance baselines and detect regressions across browser Core Web Vitals (LCP, INP, CLS, page weight), API endpoint latency percentiles, and build/test feedback times, with before/after…

    276k GitHub starsUsed in 3 repos~654 tokens
    Frontend & DesignAuto-check passed
  • Benchmark

    affaan-m/ECC

    このスキルを使用して、パフォーマンスベースラインを測定し、PR前後の回帰を検出し、スタック代替案を比較します. An agent skill from affaan-m/ECC.

    276k GitHub stars~412 tokensUpdated 4 days ago
    Auto-check passed
  • Benchmark

    affaan-m/ECC

    使用此技能测量性能基线,检测PR前后的回归,并比较堆栈替代方案。

    276k GitHub stars~330 tokensUpdated 4 days ago
    Auto-check passed
  • Cost Tracking

    affaan-m/ECC

    Track and report Claude Code token usage, spending, and budgets from the local ECC cost-tracker metrics log.

    276k GitHub starsUsed in 1 repo~1.3k tokens
    AI & LLM EngineeringAuto-check passed
  • Establishes page load, Core Web Vitals and resource-size baselines, then compares before and after on every pull request to track performance trends over time.

    136k GitHub stars~7.3k tokensUpdated today
    Frontend & DesignAuto-check: notes
  • Cross-Model Benchmark

    garrytan/gstack

    Sends one prompt to Claude, GPT through the Codex CLI and Gemini, then tabulates response time, token use and cost, with an optional judged quality score.

    136k GitHub stars~4k tokensUpdated today
    AI & LLM EngineeringAuto-check: notes

More from ruvnet/ruflo

All 264 skills in this repo
  • Stores, searches, and retrieves successful patterns with HNSW-indexed semantic search so agents can reuse past solutions instead of relearning them.

    74k GitHub starsUsed in 2 repos~830 tokens
    Auto-check passed
  • Runs claude-flow CLI security scans for input validation, path traversal, SQL injection, XSS, hardcoded secrets and known CVEs, and writes an audit report.

    74k GitHub starsUsed in 2 repos~823 tokens
    Auto-check passed
  • Applies the SPARC method (specification, pseudocode, architecture, refinement, completion) with 17 specialized modes and multi-agent orchestration, from research to deployment.

    74k GitHub starsUsed in 2 repos~829 tokens
    Auto-check passed
  • Coordinates a hierarchical swarm of specialized agents through the claude-flow CLI for work that spans several files or modules at once.

    74k GitHub starsUsed in 2 repos~779 tokens
    Auto-check passed
  • Sets up and drives Ruflo, an npm-installed orchestration layer for multi-agent swarms, persistent memory, routing, hooks and its MCP tool catalog.

    74k GitHub starsUsed in 1 repo~975 tokens
    Auto-check passed
  • Agent Coordination

    ruvnet/ruflo

    Reference for spawning, listing, monitoring and stopping agents with claude-flow commands, with agent type families, routing codes and coordination tips.

    74k GitHub starsUsed in 2 repos~519 tokens
    Auto-check passed

Questions about Cost Trend

What does Cost Trend do?

Read every docs/benchmarks/runs/.json and surface drift in win rate, latency, escalation rate, and LLM-baseline cost over time. Cost Trend is an agent skill from ruvnet/ruflo.

How do I install Cost Trend in Claude Code?

Run `npx skills add ruvnet/ruflo --skill cost-trend -a claude-code`. Or copy the skill folder (plugins/ruflo-cost-tracker/skills/cost-trend in ruvnet/ruflo) into .claude/skills/cost-trend in your project. Claude Code loads it when a task matches its description.

How do I install Cost Trend in Codex?

Run `npx skills add ruvnet/ruflo --skill cost-trend -a codex`. Or copy the skill folder (plugins/ruflo-cost-tracker/skills/cost-trend in ruvnet/ruflo) into .agents/skills/cost-trend in your project. Codex loads it when a task matches its description.

Can I use Cost Trend in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add ruvnet/ruflo --skill cost-trend -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/cost-trend, .gemini/skills/cost-trend, .github/skills/cost-trend and .opencode/skills/cost-trend in your project.

What does Cost Trend need to run?

Going by SKILL.md and its folder, Cost Trend needs the command-line tools its instructions call (node). Its frontmatter pre-approves these tools: Bash.

Does Cost Trend access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is Cost Trend safe to install?

Our automated static check of SKILL.md found notes only (pre-approves every shell command (allowed-tools: bash)), nothing it rates as a warning. It is not a guarantee. Review the folder before installing.

What licence does Cost Trend use?

Cost Trend is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Cost Trend use?

About 490 tokens (SKILL.md is roughly 2k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Cost Trend?

Skills that share tags, products or a category with Cost Trend: Benchmark (affaan-m/ECC, 276k stars), Benchmark (affaan-m/ECC, 276k stars), Benchmark (affaan-m/ECC, 276k stars) and Cost Tracking (affaan-m/ECC, 276k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Cost Trend?

ruvnet (a GitHub user) maintains it in ruvnet/ruflo, which has 74,159 GitHub stars. The repository holds 264 skills in this directory. The repository was last updated on October 9, 2026.

Source: ruvnet/ruflo on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.