Topic · AI & LLM Engineering

Best LLM cost and token optimization skills, page 3

Skills #97–144 of 259, ranked by score.

LLM cost and token optimization skills, ranked

Ranked by score. Sort bymost stars,trending,newest,recently updated

LLM cost and token optimization skills, ranked
#SkillRepositoryStarsUsed inTokensAuto-checkLicenceUpdated
97

This skill should be used when creating conversation summaries for CodeMemory, discussing summarization quality, optimizing DAG structure, or working with the compaction system.

harrylettering/CodeMemory158—~788Automated safety check: PassNo licence2 days ago
98

Manage context engineering configurations, RTK filter sets, and conversation sessions from the CLI. Apply context-relay settings and inspect active…

diegosouzapw/OmniRoute74k—~1.2kAutomated safety check: PassMITtoday
99

Configures OmniRoute's RTK, Caveman and stacked compression modes, manages language packs and rules, and previews token savings of 60 to 90 percent.

diegosouzapw/OmniRoute74k—~1.6kAutomated safety check: PassMITtoday
100

Verifies how Cline marks Anthropic prompt-cache breakpoints on the wire, then adds a tools breakpoint and tunes the cache TTL in its AI SDK code.

OnlyTerp/prompt-cache-skills114—~2kAutomated safety check: PassUnknown1 mo ago
101

Configure token budget limits, auto-compact settings, and view current budget status (model-aware — Claude 5 lineup, Opus 5.5 default fallback, full 1M context at standard price)

egorfedorov/claude-context-optimizer114—~808Automated safety check: PassMIT8 days ago
102

Builds an explorable HTML report of Claude Code usage over a chosen period: tokens, cache behavior, subagents, skills and costly prompts.

anthropics/claude-plugins-official38k—~784Automated safety check: PassApache-2.0today
103

Always-on context compression for OpenClaw. An agent skill from Supercompress/Supercompress.

Supercompress/Supercompress106—~366Automated safety check: PassMITyesterday
104

Estimate Kubernetes infrastructure costs by querying cluster node, pod, PVC/PV, and LoadBalancer data, applying cloud pricing models, and producing cost attribution reports with storage and…

initializ/forge222—~2.7kAutomated safety check: PassApache-2.06 days ago
105

Patch guide for adding a stable per-task prompt_cache_key to Cline's OpenAI native provider so cached token counts stop reading as zero.

OnlyTerp/prompt-cache-skills114—~868Automated safety check: PassUnknown1 mo ago
106

This skill should be used for improving context efficiency: context budgeting, observation masking, prefix or KV-cache strategy, partitioning, token-cost reduction, retrieval scoping, and extending…

guanyang/open-agent-hub9752 repos~4kAutomated safety check: PassMITyesterday
107

This skill should be used for persistent semantic memory in agent systems: cross-session knowledge retention, entity tracking, temporal validity, graph or vector retrieval, memory consolidation, and…

guanyang/open-agent-hub9752 repos~4.1kAutomated safety check: PassMITyesterday
108

An advanced skill for L3 autonomous loops. An agent skill from cobusgreyling/loop-engineering.

cobusgreyling/loop-engineering11k1 repo~864Automated safety check: PassMITtoday
109

Add a new provider API capability (prompt caching, strict/structured tool calling, thinking/reasoning effort, service tier, safety settings, logprobs, etc.) to Pydantic AI.

pydantic/pydantic-ai20k—~3.2kAutomated safety check: PassMITtoday
110

Prepares many-file, single-turn transforms such as translating or rewriting as a plan, then submits it to the asynchronous, half-price DashScope Batch API through the qwen batch CLI.

QwenLM/qwen-code28k—~2.2kAutomated safety check: PassApache-2.0today
111

Standardized cost-estimation framework for greatcto plans. An agent skill from avelikiy/great_cto.

avelikiy/great_cto1031 repo~1.3kAutomated safety check: PassMITtoday
112

Analyze prompt-cache effectiveness for Claude Code usage from the Agent Monitor dashboard — cache hit rate (totalcacheread / (totalcacheread + totalinput)), cachewrite vs cacheread reuse, cache-read…

hoangsonww/Claude-Code-Agent-Monitor1.1k—~966Automated safety check: PassMITyesterday
113

Caching strategies for LLM prompts including Anthropic prompt caching, response caching, and CAG (Cache Augmented Generation) Use when: prompt caching, cache prompt, response cache, cag, cache…

davila7/claude-code-templates32k6 repos~452Automated safety check: PassMITtoday
114

Cache the parts of the prompt that don't change so a long-running loop stops paying full price on every turn.

Archive228/loopkit756—~735Automated safety check: PassMIT2 mo ago
115

A skill your agent uses when the user wants to capture AI tool costs, review spending trends, set cost budgets, or integrate cost data into health snapshots — guides quarterly cost capture, records…

Habitat-Thinking/ai-literacy-superpowers114—~1.7kAutomated safety check: PassUnknown17 days ago
116

Specifies how quota-axi reads LLM subscription quota windows, derives pace and runway, computes a selection signal and renders output in TOON, JSON and TUI tiers.

kunchenguid/quota-axi144—~1.2kAutomated safety check: PassMITtoday
117

Guide for debugging OhMyCode issues. An agent skill from AlphaLab-USTC/OhMyCode.

AlphaLab-USTC/OhMyCode131—~1.3kAutomated safety check: PassMIT6 mo ago
118

Checks every subagent prompt before spawning, swapping pasted files and context for paths and short summaries and trimming the brief, to avoid multiplied token cost.

LichAmnesia/lich-skills234—~1.6kAutomated safety check: PassMIT4 mo ago
119

Track and report Claude Code token usage, spending, and budgets from the local ECC cost-tracker metrics log.

affaan-m/ECC275k1 repo~1.3kAutomated safety check: PassMIT3 days ago
120

Continue's prompt caching is opt-in via config and off by default.

OnlyTerp/prompt-cache-skills114—~977Automated safety check: PassUnknown1 mo ago
121

A skill your agent uses when the user asks to optimize prompts, design prompt templates, evaluate LLM outputs with an eval set, measure RAG retrieval quality, validate agent/tool configurations…

alirezarezvani/claude-skills28k1 repo~2.5kAutomated safety check: PassMIT1 mo ago
122

Generate a concise daily infrastructure briefing. An agent skill from bolivian-peru/os-moda.

bolivian-peru/os-moda119—~897Automated safety check: PassApache-2.03 mo ago
123

A skill your agent uses to classify the documentation impact of a pull request diff, returning one of three verdicts -- no-change, in-place edit, or structural change -- with bounded LLM cost.

microsoft/apm4k—~1.8kAutomated safety check: PassMITtoday
124

This skill should be used when writing, enhancing, or evaluating the launch prompt for a long-running autonomous agent or a parallel multi-agent orchestration attacking a hard problem: pseudo-formal…

guanyang/open-agent-hub9751 repo~6.4kAutomated safety check: PassMITyesterday
125

This skill should be used when a model gets read-write control over its own live context window instead of a harness-scheduled compaction policy: the context exposed as an editable file the model…

guanyang/open-agent-hub9751 repo~6kAutomated safety check: PassMITyesterday
126
126.Cost

AI Build Cost Tracker — track how much AI is costing you per feature.

Houseofmvps/ultraship123—~1.1kAutomated safety check: NotesMIT3 mo ago
127

Orchestrate free opencode workers from Claude Code to cut token costs.

alirezarezvani/claude-skills28k1 repo~2.4kAutomated safety check: PassApache-2.01 mo ago
128

A skill your agent uses when context compression is imminent, when resuming a session, or when preserving critical decisions across long tasks.

vibeeval/vibecosystem531—~1.7kAutomated safety check: PassMIT2 mo ago
129

Context window management and token optimization guide. An agent skill from jh941213/my-cc-harness.

jh941213/my-cc-harness126—~621Automated safety check: PassNo licence2 mo ago
130

Track session costs, set budget alerts, and optimize token spend.

rohitg00/pro-workflow2.9k—~814Automated safety check: PassNo licence8 days ago
131

Reduce token waste by 40-60% through anti-sycophancy rules, tool-call budgets, one-pass coding, task profiles, and read-before-write enforcement.

rohitg00/pro-workflow2.9k1 repo~1kAutomated safety check: PassNo licence8 days ago
132

Analyze markdown files for token efficiency and reduce context-window bloat.

Azure/azure-sdk-tools134—~602Automated safety check: PassMITtoday
133

Analyze and reduce token consumption in agentic workflows — guardrail-specific entry points, measurement, and optimization techniques.

github/gh-aw5.4k—~1.2kAutomated safety check: PassMITtoday
134
134.App ObservabilityOfficial

Get RED metrics + service maps + frontend RUM + AI/LLM monitoring out of Grafana Cloud — Application Observability (tracesspanmetrics from OTel traces, p50/p95/p99 latency, exemplar-to-trace…

grafana/skills279—~1.8kAutomated safety check: PassApache-2.0yesterday
135

Monitor LLMs and agentic apps: performance, token/cost, response quality, and workflow orchestration.

aspectrr/deer405—~858Automated safety check: PassMIT5 mo ago
136

Create and manage webhook subscriptions for event-driven agent activation, or for direct push notifications (zero LLM cost).

Tommy-yw/RunbookHermes5461 repo~1.7kAutomated safety check: NotesMIT4 mo ago
137

追蹤並計算當前 session 的 Token 使用量和花費(USD)。產出前後對比報告,幫助使用者了解 workspace 優化的實際效益。

zeuikli/claude-code-workspace156—~830Automated safety check: PassNo licenceyesterday
138

Orchestrate autonomous AI development pipelines through your Kanban board (Asana, GitHub Projects, Linear).

sickn33/agentic-awesome-skills47k2 repos~2kAutomated safety check: PassMITtoday
139

Export cost-tracking telemetry in Prometheus textfile or webhook JSON formats — for external observability (Grafana, Datadog, custom dashboards)

ruvnet/ruflo74k1 repo~687Automated safety check: NotesMITtoday
140

Analyze token usage patterns and recommend cost optimizations with estimated savings

ruvnet/ruflo74k1 repo~997Automated safety check: NotesMITtoday
141

Generate a cost report showing token usage and USD costs by agent and model

ruvnet/ruflo74k1 repo~830Automated safety check: NotesMITtoday
142

Single-shot programmatic dump of all cost data — total spend, per-tier, top session, budget status, federation aggregate.

ruvnet/ruflo74k1 repo~594Automated safety check: NotesMITtoday
143

OpenCode places cachePoint on Bedrock messages containing DocumentBlocks, which produces a "nothing available to cache" error.

OnlyTerp/prompt-cache-skills114—~966Automated safety check: PassUnknown1 mo ago
144

A skill your agent uses when managing prompts in production at scale: versioning prompts, running A/B tests on prompts, building prompt registries, preventing prompt regressions, or creating eval…

alirezarezvani/claude-skills28k1 repo~2.8kAutomated safety check: PassMIT1 mo ago