Agent skill

Model Savings

by hoangsonww in hoangsonww/Claude-Code-Agent-Monitor

Estimate the dollars saved by routing eligible Claude Code work to a cheaper model family, using the Agent Monitor pricing engine.

MITAuto-check passed

Install Model Savings

skills CLI
$ npx skills add hoangsonww/Claude-Code-Agent-Monitor --skill model-savings -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install hoangsonww/Claude-Code-Agent-Monitor model-savings --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/hoangsonww/Claude-Code-Agent-Monitor.git skills-src && mkdir -p .claude/skills && cp -r skills-src/plugins/ccam-cost-guard/skills/model-savings .claude/skills/model-savings && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
model-savings
GitHub stars
1.1k
Token cost
~1k tokens
SKILL.md length
418 words
Files
2
Skills in repo
78
Repo updated
First seen
Licence
MIT

At a glance

Estimate the dollars saved by routing eligible Claude Code work to a cheaper model family, using the Agent Monitor pricing engine.

  • Works in 5 steps: Current spend by model → Re-priced at target family → Eligible-only estimate → …
  • Hunting for cost cuts
  • SKILL.md covers Input, Data Sources, Savings method and Report Sections, plus 1 more section
  • Instructions only: no scripts, shell commands, URLs or credentials in SKILL.md

What it does

Model Savings is an agent skill from hoangsonww/Claude-Code-Agent-Monitor. Estimate the dollars saved by routing eligible Claude Code work to a cheaper model family, using the Agent Monitor pricing engine. Re-prices each model's token mix at the target family's rates and quantifies the delta. Uses /api/pricing (rates), /api/pricing/cost (current per-model spend), /api/sessions, and /api/analytics. Use when hunting for cost cuts or comparing model tiers.

Its SKILL.md is about 1k tokens, which your agent loads only when the skill is triggered. The skill folder holds 2 other files (for example `agents/openai.yaml`).

The repository describes itself as: 🚀 A real-time monitoring dashboard for Claude Code & Codex, built with SQLite3, Node.js, Express, React, Vite, TailwindCSS, & WebSockets. It tracks sessions, agent activity… The licence is MIT.

When your agent uses it

  • Hunting for cost cuts
  • Comparing model tiers

Example prompts

  • “s token mix at the target family”
  • “/model-savings”

Workflow steps

5 steps, taken from the step headings in SKILL.md.

  1. Current spend by model
  2. Re-priced at target family
  3. Eligible-only estimate
  4. Recommended routing
  5. Caveats

What it can do on your machine

Read from SKILL.md and the folder at commit a06db03. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    No scripts in the folder and no shell commands in SKILL.md.

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Model Savings loads about 1k tokens when it runs. Until then it costs about 99 tokens; SKILL.md has 418 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~99
When it runs · the whole SKILL.md, loaded when a task matches
~1k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from hoangsonww/Claude-Code-Agent-Monitor at commit a06db03, republished under its MIT licence (© hoangsonww). 418 words, ~1,000 tokens.

Download SKILL.mdSave it as .claude/skills/model-savings/SKILL.md (or your agent's skills folder). This skill also uses 1 other file; get the full folder from GitHub.
name
model-savings
description
Estimate the dollars saved by routing eligible Claude Code work to a cheaper model family, using the Agent Monitor pricing engine. Re-prices each model's token mix at the target family's rates and quantifies the delta. Uses /api/pricing (rates), /api/pricing/cost (current per-model spend), /api/sessions, and /api/analytics. Use when hunting for cost cuts or comparing model tiers.

Model Savings

Quantify how much spend you would recover by moving eligible work to a cheaper model.

Input

The user provides: $ARGUMENTS

This is the routing question — e.g. "Opus → Sonnet", "move simple work to Haiku", or empty (analyze every premium model against the next tier down). If no target family is named, default to proposing the next-cheaper tier per model and say so.

Data Sources

EndpointReturns
GET /api/pricing{ pricing: [{ model_pattern, display_name, input_per_mtok, output_per_mtok, cache_read_per_mtok, cache_write_per_mtok }] } — the rate card for every family
GET /api/pricing/cost{ total_cost, breakdown: [{ model, input_tokens, output_tokens, cache_read_tokens, cache_write_tokens, cost, matched_rule }] } — current spend and the exact token mix per model
GET /api/sessions?limit=200Sessions with model, inline cost, and metadata (turn_count, thinking_blocks) — used to judge which work is eligible to downshift
GET /api/analyticsagent_types, tool_usage, total_subagents — corroborate which task types are low-complexity and safe to route cheaper

Savings method

For each candidate model in the cost breakdown, re-price its exact token mix at the target family's rates:

cost_at_target = (input_tokens      / 1M) × target.input_per_mtok
               + (output_tokens     / 1M) × target.output_per_mtok
               + (cache_read_tokens / 1M) × target.cache_read_per_mtok
               + (cache_write_tokens/ 1M) × target.cache_write_per_mtok

savings = current_model_cost − cost_at_target

Pull target.*_per_mtok from /api/pricing (longest model_pattern match wins). Default rates ($/Mtok in/out/cacheRead/cacheWrite): Opus $5/$25/$0.50/$6.25, Sonnet $3/$15/$0.30/$3.75, Haiku $1/$5/$0.10/$1.25.

Eligibility — don't promise savings on work that needs the big model

Re-pricing the full token mix is the theoretical ceiling. Scope it to eligible work:

  • Low-turn sessions (metadata.turn_count small) and simple subagent/tool work are safe to downshift.
  • Heavy-reasoning sessions (many thinking_blocks, high turn counts) likely need the premium model — exclude or discount them.
  • Report both the full re-price (ceiling) and an eligible-only estimate, and state the eligibility rule you applied.

Report Sections

Show full SKILL.md (170 more words)Show less
1. Current spend by model

Table from /api/pricing/cost: each model, its 4 token counts, and current cost. Note its share of total_cost.

2. Re-priced at target family

For each candidate, show cost_at_target and savings (absolute $ and %). Make the target rate card explicit.

3. Eligible-only estimate

Apply the eligibility rule and recompute savings over just the downshiftable token mix. Show how many sessions / what share of tokens qualified.

Rank routing moves by eligible monthly savings (descending), top 5. For each: source → target, the token mix moved, estimated $ saved, and a confidence level (high/medium/low) based on how clearly the work is low-complexity.

5. Caveats

Cheaper models may need more turns or produce more output — note that realized savings can be lower than the static re-price, and that quality-sensitive work should stay on the premium tier.

Output

Markdown tables. Currency as USD to 4 decimal places; token counts with thousands separators; rates as $/Mtok. Always present both the ceiling (full re-price) and the eligible-only estimate so the number is honest.

© hoangsonww, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

SKILL.md and 1 other file in plugins/ccam-cost-guard/skills/model-savings of hoangsonww/Claude-Code-Agent-Monitor.

  • SKILL.md
  • agents/openai.yaml

Open the folder on GitHubat commit a06db03

Compare with similar skills

Model Savings next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Model Savings compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Model Savings this skillhoangsonww/Claude-Code-Agent-Monitor1.1k—~1kAutomated safety check: PassMIT
Task Effort EstimatorDonchitos/Claude-Code-Game-Studios26k—~1.2kAutomated safety check: PassMIT
Progressive Estimationsickn33/agentic-awesome-skills47k2 repos~863Automated safety check: PassMIT
Context Savegarrytan/gstack136k—~9.9kAutomated safety check: NotesMIT
Save As PDFasgeirtj/system_prompts_leaks69k—~2.3kAutomated safety check: PassCC0-1.0
Save to Obsidian VaultAgriciDaniel/claude-obsidian15k—~1.3kAutomated safety check: PassMIT

Similar skills

  • Task Effort Estimator

    Donchitos/Claude-Code-Game-Studios

    Estimates the effort for a game development task from code complexity, scope, risk and past sprint data, returning a range with a confidence level.

    26k GitHub stars~1.2k tokensUpdated 3 days ago
    Product & Project ManagementAuto-check passed
  • Progressive Estimation

    sickn33/agentic-awesome-skills

    Estimate AI-assisted and hybrid human+agent development work with research-backed PERT statistics and calibration feedback loops

    47k GitHub starsUsed in 2 repos~863 tokens
    Data & AnalyticsAuto-check passed
  • Context Save

    garrytan/gstack

    Captures git state, decisions made and remaining work so that a later session can resume the task without losing context.

    136k GitHub stars~9.9k tokensUpdated today
    Agent WorkflowsAuto-check: notes
  • Save As PDF

    asgeirtj/system_prompts_leaks

    Print-ready PDF export

    69k GitHub stars~2.3k tokensUpdated yesterday
    Documents & OfficeAuto-check passed
  • Save to Obsidian Vault

    AgriciDaniel/claude-obsidian

    Files a specific answer, decision or insight you point to in a conversation as one reviewed note in your Obsidian vault, and only when you ask for it.

    15k GitHub stars~1.3k tokensUpdated 1 mo ago
    Knowledge ManagementAuto-check passed
  • Official

    Encrypted Saved Objects (ESO) in Kibana — registration, AAD attribute choices, partial update safety, model version migrations with createModelVersion, canEncrypt checks, and Serverless constraints.

    21k GitHub stars~4.3k tokensUpdated today
    Backend & APIsAuto-check passed

More from hoangsonww/Claude-Code-Agent-Monitor

All 78 skills in this repo
  • Version Release

    hoangsonww/Claude-Code-Agent-Monitor

    Choose and apply the correct semantic version bump for this repository.

    1.1k GitHub stars~1.6k tokensUpdated yesterday
    Auto-check passed
  • Budget Set

    hoangsonww/Claude-Code-Agent-Monitor

    Define a spend budget for Claude Code and, optionally, create a cost alert rule that fires when usage crosses the limit, via POST /api/alerts/rules on the Agent Monitor dashboard.

    1.1k GitHub stars~1k tokensUpdated yesterday
    Auto-check passed
  • Cache Efficiency

    hoangsonww/Claude-Code-Agent-Monitor

    Analyze prompt-cache effectiveness for Claude Code usage from the Agent Monitor dashboard — cache hit rate (totalcacheread / (totalcacheread + totalinput)), cachewrite vs cacheread reuse, cache-read…

    1.1k GitHub stars~966 tokensUpdated yesterday
    Auto-check passed
  • Cost Breakdown

    hoangsonww/Claude-Code-Agent-Monitor

    Break down Claude Code costs using the Agent Monitor pricing engine.

    1.1k GitHub stars~845 tokensUpdated yesterday
    Auto-check passed
  • Dag Map

    hoangsonww/Claude-Code-Agent-Monitor

    Render the multi-agent orchestration DAG for a session — parent→child subagent edges, tree depth, and fan-out — from the Agent Monitor workflow intelligence API.

    1.1k GitHub stars~564 tokensUpdated yesterday
    Auto-check passed
  • Dashboard Status

    hoangsonww/Claude-Code-Agent-Monitor

    Quick dashboard health and status overview — checks the Agent Monitor API (port 4820), reports session/agent/event counts from /api/stats, confirms WebSocket connectivity, reads the redacted hook…

    1.1k GitHub stars~600 tokensUpdated yesterday
    Auto-check passed

Questions about Model Savings

What does Model Savings do?

Estimate the dollars saved by routing eligible Claude Code work to a cheaper model family, using the Agent Monitor pricing engine. Model Savings is an agent skill from hoangsonww/Claude-Code-Agent-Monitor. Estimate the dollars saved by routing eligible Claude Code work to a cheaper model family, using the Agent Monitor pricing engine.

When should I use Model Savings?

Model Savings fits situations like: hunting for cost cuts; comparing model tiers.

How do I install Model Savings in Claude Code?

Run `npx skills add hoangsonww/Claude-Code-Agent-Monitor --skill model-savings -a claude-code`. Or copy the skill folder (plugins/ccam-cost-guard/skills/model-savings in hoangsonww/Claude-Code-Agent-Monitor) into .claude/skills/model-savings in your project. Claude Code loads it when a task matches its description.

How do I install Model Savings in Codex?

Run `npx skills add hoangsonww/Claude-Code-Agent-Monitor --skill model-savings -a codex`. Or copy the skill folder (plugins/ccam-cost-guard/skills/model-savings in hoangsonww/Claude-Code-Agent-Monitor) into .agents/skills/model-savings in your project. Codex loads it when a task matches its description.

Can I use Model Savings in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add hoangsonww/Claude-Code-Agent-Monitor --skill model-savings -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/model-savings, .gemini/skills/model-savings, .github/skills/model-savings and .opencode/skills/model-savings in your project.

What does Model Savings need to run?

SKILL.md names no scripts, command-line tools or credentials: Model Savings is instructions for the agent only.

Does Model Savings access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is Model Savings safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Model Savings use?

Model Savings is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Model Savings use?

About 1k tokens (SKILL.md is roughly 4k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Model Savings?

Skills that share tags, products or a category with Model Savings: Task Effort Estimator (Donchitos/Claude-Code-Game-Studios, 26k stars), Progressive Estimation (sickn33/agentic-awesome-skills, 47k stars), Context Save (garrytan/gstack, 136k stars) and Save As PDF (asgeirtj/system_prompts_leaks, 69k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Model Savings?

hoangsonww (a GitHub user) maintains it in hoangsonww/Claude-Code-Agent-Monitor, which has 1,058 GitHub stars. The repository holds 78 skills in this directory. The repository was last updated on October 10, 2026.

Source: hoangsonww/Claude-Code-Agent-Monitor on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.