Topic · AI & LLM Engineering
Best LLM cost and token optimization skills, page 3
LLM cost and token optimization skills, ranked
Ranked by score. Sort bymost stars,trending,newest,recently updated
| # | Skill | Repository | Stars | Used in | Tokens | Auto-check | Licence | Updated |
|---|---|---|---|---|---|---|---|---|
| 97 | This skill should be used when creating conversation summaries for CodeMemory, discussing summarization quality, optimizing DAG structure, or working with the compaction system. | harrylettering/ | 158 | — | ~788 | Automated safety check: Pass | No licence | 2 days ago |
| 98 | Manage context engineering configurations, RTK filter sets, and conversation sessions from the CLI. Apply context-relay settings and inspect active… | diegosouzapw/ | 74k | — | ~1.2k | Automated safety check: Pass | MIT | today |
| 99 | Configures OmniRoute's RTK, Caveman and stacked compression modes, manages language packs and rules, and previews token savings of 60 to 90 percent. | diegosouzapw/ | 74k | — | ~1.6k | Automated safety check: Pass | MIT | today |
| 100 | Verifies how Cline marks Anthropic prompt-cache breakpoints on the wire, then adds a tools breakpoint and tunes the cache TTL in its AI SDK code. | OnlyTerp/ | 114 | — | ~2k | Automated safety check: Pass | Unknown | 1 mo ago |
| 101 | 101.Cco Budget Configure token budget limits, auto-compact settings, and view current budget status (model-aware — Claude 5 lineup, Opus 5.5 default fallback, full 1M context at standard price) | egorfedorov/ | 114 | — | ~808 | Automated safety check: Pass | MIT | 8 days ago |
| 102 | Builds an explorable HTML report of Claude Code usage over a chosen period: tokens, cache behavior, subagents, skills and costly prompts. | anthropics/ | 38k | — | ~784 | Automated safety check: Pass | Apache-2.0 | today |
| 103 | 103.Supercompress Always-on context compression for OpenClaw. An agent skill from Supercompress/Supercompress. | Supercompress/ | 106 | — | ~366 | Automated safety check: Pass | MIT | yesterday |
| 104 | Estimate Kubernetes infrastructure costs by querying cluster node, pod, PVC/PV, and LoadBalancer data, applying cloud pricing models, and producing cost attribution reports with storage and… | initializ/ | 222 | — | ~2.7k | Automated safety check: Pass | Apache-2.0 | 6 days ago |
| 105 | Patch guide for adding a stable per-task prompt_cache_key to Cline's OpenAI native provider so cached token counts stop reading as zero. | OnlyTerp/ | 114 | — | ~868 | Automated safety check: Pass | Unknown | 1 mo ago |
| 106 | This skill should be used for improving context efficiency: context budgeting, observation masking, prefix or KV-cache strategy, partitioning, token-cost reduction, retrieval scoping, and extending… | guanyang/ | 975 | 2 repos | ~4k | Automated safety check: Pass | MIT | yesterday |
| 107 | 107.Memory Systems This skill should be used for persistent semantic memory in agent systems: cross-session knowledge retention, entity tracking, temporal validity, graph or vector retrieval, memory consolidation, and… | guanyang/ | 975 | 2 repos | ~4.1k | Automated safety check: Pass | MIT | yesterday |
| 108 | An advanced skill for L3 autonomous loops. An agent skill from cobusgreyling/loop-engineering. | cobusgreyling/ | 11k | 1 repo | ~864 | Automated safety check: Pass | MIT | today |
| 109 | Add a new provider API capability (prompt caching, strict/structured tool calling, thinking/reasoning effort, service tier, safety settings, logprobs, etc.) to Pydantic AI. | pydantic/ | 20k | — | ~3.2k | Automated safety check: Pass | MIT | today |
| 110 | Prepares many-file, single-turn transforms such as translating or rewriting as a plan, then submits it to the asynchronous, half-price DashScope Batch API through the qwen batch CLI. | QwenLM/ | 28k | — | ~2.2k | Automated safety check: Pass | Apache-2.0 | today |
| 111 | 111.Cost Model Standardized cost-estimation framework for greatcto plans. An agent skill from avelikiy/great_cto. | avelikiy/ | 103 | 1 repo | ~1.3k | Automated safety check: Pass | MIT | today |
| 112 | 112.Cache Efficiency Analyze prompt-cache effectiveness for Claude Code usage from the Agent Monitor dashboard — cache hit rate (totalcacheread / (totalcacheread + totalinput)), cachewrite vs cacheread reuse, cache-read… | hoangsonww/ | 1.1k | — | ~966 | Automated safety check: Pass | MIT | yesterday |
| 113 | 113.Prompt Caching Caching strategies for LLM prompts including Anthropic prompt caching, response caching, and CAG (Cache Augmented Generation) Use when: prompt caching, cache prompt, response cache, cag, cache… | davila7/ | 32k | 6 repos | ~452 | Automated safety check: Pass | MIT | today |
| 114 | 114.Prompt Caching Cache the parts of the prompt that don't change so a long-running loop stops paying full price on every turn. | Archive228/ | 756 | — | ~735 | Automated safety check: Pass | MIT | 2 mo ago |
| 115 | 115.Cost Tracking A skill your agent uses when the user wants to capture AI tool costs, review spending trends, set cost budgets, or integrate cost data into health snapshots — guides quarterly cost capture, records… | Habitat-Thinking/ | 114 | — | ~1.7k | Automated safety check: Pass | Unknown | 17 days ago |
| 116 | Specifies how quota-axi reads LLM subscription quota windows, derives pace and runway, computes a selection signal and renders output in TOON, JSON and TUI tiers. | kunchenguid/ | 144 | — | ~1.2k | Automated safety check: Pass | MIT | today |
| 117 | 117.Debug Ohmycode Guide for debugging OhMyCode issues. An agent skill from AlphaLab-USTC/OhMyCode. | AlphaLab-USTC/ | 131 | — | ~1.3k | Automated safety check: Pass | MIT | 6 mo ago |
| 118 | Checks every subagent prompt before spawning, swapping pasted files and context for paths and short summaries and trimming the brief, to avoid multiplied token cost. | LichAmnesia/ | 234 | — | ~1.6k | Automated safety check: Pass | MIT | 4 mo ago |
| 119 | 119.Cost Tracking Track and report Claude Code token usage, spending, and budgets from the local ECC cost-tracker metrics log. | affaan-m/ | 275k | 1 repo | ~1.3k | Automated safety check: Pass | MIT | 3 days ago |
| 120 | Continue's prompt caching is opt-in via config and off by default. | OnlyTerp/ | 114 | — | ~977 | Automated safety check: Pass | Unknown | 1 mo ago |
| 121 | A skill your agent uses when the user asks to optimize prompts, design prompt templates, evaluate LLM outputs with an eval set, measure RAG retrieval quality, validate agent/tool configurations… | alirezarezvani/ | 28k | 1 repo | ~2.5k | Automated safety check: Pass | MIT | 1 mo ago |
| 122 | 122.Morning Briefing Generate a concise daily infrastructure briefing. An agent skill from bolivian-peru/os-moda. | bolivian-peru/ | 119 | — | ~897 | Automated safety check: Pass | Apache-2.0 | 3 mo ago |
| 123 | A skill your agent uses to classify the documentation impact of a pull request diff, returning one of three verdicts -- no-change, in-place edit, or structural change -- with bounded LLM cost. | microsoft/ | 4k | — | ~1.8k | Automated safety check: Pass | MIT | today |
| 124 | This skill should be used when writing, enhancing, or evaluating the launch prompt for a long-running autonomous agent or a parallel multi-agent orchestration attacking a hard problem: pseudo-formal… | guanyang/ | 975 | 1 repo | ~6.4k | Automated safety check: Pass | MIT | yesterday |
| 125 | This skill should be used when a model gets read-write control over its own live context window instead of a harness-scheduled compaction policy: the context exposed as an editable file the model… | guanyang/ | 975 | 1 repo | ~6k | Automated safety check: Pass | MIT | yesterday |
| 126 | 126.Cost AI Build Cost Tracker — track how much AI is costing you per feature. | Houseofmvps/ | 123 | — | ~1.1k | Automated safety check: Notes | MIT | 3 mo ago |
| 127 | 127.Hivemind Orchestrate free opencode workers from Claude Code to cut token costs. | alirezarezvani/ | 28k | 1 repo | ~2.4k | Automated safety check: Pass | Apache-2.0 | 1 mo ago |
| 128 | 128.Notepad System A skill your agent uses when context compression is imminent, when resuming a session, or when preserving critical decisions across long tasks. | vibeeval/ | 531 | — | ~1.7k | Automated safety check: Pass | MIT | 2 mo ago |
| 129 | 129.Compact Guide Context window management and token optimization guide. An agent skill from jh941213/my-cc-harness. | jh941213/ | 126 | — | ~621 | Automated safety check: Pass | No licence | 2 mo ago |
| 130 | 130.Cost Tracker Track session costs, set budget alerts, and optimize token spend. | rohitg00/ | 2.9k | — | ~814 | Automated safety check: Pass | No licence | 8 days ago |
| 131 | 131.Token Efficiency Reduce token waste by 40-60% through anti-sycophancy rules, tool-call budgets, one-pass coding, task profiles, and read-before-write enforcement. | rohitg00/ | 2.9k | 1 repo | ~1k | Automated safety check: Pass | No licence | 8 days ago |
| 132 | Analyze markdown files for token efficiency and reduce context-window bloat. | Azure/ | 134 | — | ~602 | Automated safety check: Pass | MIT | today |
| 133 | Analyze and reduce token consumption in agentic workflows — guardrail-specific entry points, measurement, and optimization techniques. | github/ | 5.4k | — | ~1.2k | Automated safety check: Pass | MIT | today |
| 134 | Get RED metrics + service maps + frontend RUM + AI/LLM monitoring out of Grafana Cloud — Application Observability (tracesspanmetrics from OTel traces, p50/p95/p99 latency, exemplar-to-trace… | grafana/ | 279 | — | ~1.8k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 135 | Monitor LLMs and agentic apps: performance, token/cost, response quality, and workflow orchestration. | aspectrr/ | 405 | — | ~858 | Automated safety check: Pass | MIT | 5 mo ago |
| 136 | Create and manage webhook subscriptions for event-driven agent activation, or for direct push notifications (zero LLM cost). | Tommy-yw/ | 546 | 1 repo | ~1.7k | Automated safety check: Notes | MIT | 4 mo ago |
| 137 | 137.Cost Tracker 追蹤並計算當前 session 的 Token 使用量和花費(USD)。產出前後對比報告,幫助使用者了解 workspace 優化的實際效益。 | zeuikli/ | 156 | — | ~830 | Automated safety check: Pass | No licence | yesterday |
| 138 | 138.Agentflow Orchestrate autonomous AI development pipelines through your Kanban board (Asana, GitHub Projects, Linear). | sickn33/ | 47k | 2 repos | ~2k | Automated safety check: Pass | MIT | today |
| 139 | 139.Cost Export Export cost-tracking telemetry in Prometheus textfile or webhook JSON formats — for external observability (Grafana, Datadog, custom dashboards) | ruvnet/ | 74k | 1 repo | ~687 | Automated safety check: Notes | MIT | today |
| 140 | 140.Cost Optimize Analyze token usage patterns and recommend cost optimizations with estimated savings | ruvnet/ | 74k | 1 repo | ~997 | Automated safety check: Notes | MIT | today |
| 141 | 141.Cost Report Generate a cost report showing token usage and USD costs by agent and model | ruvnet/ | 74k | 1 repo | ~830 | Automated safety check: Notes | MIT | today |
| 142 | 142.Cost Summary Single-shot programmatic dump of all cost data — total spend, per-tier, top session, budget status, federation aggregate. | ruvnet/ | 74k | 1 repo | ~594 | Automated safety check: Notes | MIT | today |
| 143 | OpenCode places cachePoint on Bedrock messages containing DocumentBlocks, which produces a "nothing available to cache" error. | OnlyTerp/ | 114 | — | ~966 | Automated safety check: Pass | Unknown | 1 mo ago |
| 144 | A skill your agent uses when managing prompts in production at scale: versioning prompts, running A/B tests on prompts, building prompt registries, preventing prompt regressions, or creating eval… | alirezarezvani/ | 28k | 1 repo | ~2.8k | Automated safety check: Pass | MIT | 1 mo ago |
Explore related skills
Category
More topics in AI & LLM Engineering
- Building AI agents563
- Deep learning415
- Embeddings386
- LLM inference and serving372
- Prompt engineering360
- Retrieval-augmented generation358
- Fine-tuning309
- LLM evaluation308
- Speech recognition and synthesis308
- Structured output and tool calling276
- LLM API integration255
- Model routing and gateways255
- LLM observability240
- LLM guardrails221
- Computer vision203
- Model hubs and datasets180
- GPU and accelerator computing176
- Diffusion and image models166
- Natural language processing131
- Reinforcement learning66
- AI interpretability23