Search
LLM cost and token optimization
Skills
Sort:BestMost starsTrending todayTrending this weekTrending this monthNewestRecently updatedName
| # | Skill | Repository | Stars | Used in | Tokens | Auto-check | Licence | Updated |
|---|---|---|---|---|---|---|---|---|
| 97 | Controls the RTK filter set and context-handling settings in OmniRoute, with endpoints to try compression on sample text and read back retained output. | diegosouzapw/ | 75k | — | ~618 | Automated safety check: Pass | MIT | today |
| 98 | Queries OmniRoute call logs, usage history and analytics, filters them by provider, model, status or cost, and reads or sets usage budgets. | diegosouzapw/ | 75k | — | ~2k | Automated safety check: Pass | MIT | today |
| 99 | Verifies how Cline marks Anthropic prompt-cache breakpoints on the wire, then adds a tools breakpoint and tunes the cache TTL in its AI SDK code. | OnlyTerp/ | 114 | — | ~2k | Automated safety check: Pass | Unknown | 1 mo ago |
| 100 | 100.Cco Budget Configure token budget limits, auto-compact settings, and view current budget status (model-aware — Claude 5 lineup, Opus 5.5 default fallback, full 1M context at standard price) | egorfedorov/ | 114 | — | ~808 | Automated safety check: Pass | MIT | 11 days ago |
| 101 | Builds an explorable HTML report of Claude Code usage over a chosen period: tokens, cache behavior, subagents, skills and costly prompts. | anthropics/ | 38k | — | ~784 | Automated safety check: Pass | Apache-2.0 | today |
| 102 | 102.Supercompress Always-on context compression for OpenClaw. An agent skill from Supercompress/Supercompress. | Supercompress/ | 107 | — | ~366 | Automated safety check: Pass | MIT | yesterday |
| 103 | Estimate Kubernetes infrastructure costs by querying cluster node, pod, PVC/PV, and LoadBalancer data, applying cloud pricing models, and producing cost attribution reports with storage and… | initializ/ | 222 | — | ~2.7k | Automated safety check: Pass | Apache-2.0 | 2 days ago |
| 104 | Check token budget and run-log spend before and after a loop run. Enforces early exit when over budget or when there is no actionable work. | cobusgreyling/ | 11k | 1 repo | ~376 | Automated safety check: Pass | MIT | today |
| 105 | Patch guide for adding a stable per-task prompt_cache_key to Cline's OpenAI native provider so cached token counts stop reading as zero. | OnlyTerp/ | 114 | — | ~868 | Automated safety check: Pass | Unknown | 1 mo ago |
| 106 | This skill should be used for improving context efficiency: context budgeting, observation masking, prefix or KV-cache strategy, partitioning, token-cost reduction, retrieval scoping, and extending… | guanyang/ | 977 | 2 repos | ~4k | Automated safety check: Pass | MIT | today |
| 107 | 107.Memory Systems This skill should be used for persistent semantic memory in agent systems: cross-session knowledge retention, entity tracking, temporal validity, graph or vector retrieval, memory consolidation, and… | guanyang/ | 977 | 2 repos | ~4.1k | Automated safety check: Pass | MIT | today |
| 108 | 108.Prompt Engineer Designs, tests and refines LLM prompts: zero-shot, few-shot and chain-of-thought patterns, system prompts, structured output schemas and evaluation test suites. | Jeffallan/ | 12k | — | ~1.5k | Automated safety check: Pass | MIT | 7 days ago |
| 109 | An advanced skill for L3 autonomous loops. An agent skill from cobusgreyling/loop-engineering. | cobusgreyling/ | 11k | 1 repo | ~864 | Automated safety check: Pass | MIT | today |
| 110 | Add a new provider API capability (prompt caching, strict/structured tool calling, thinking/reasoning effort, service tier, safety settings, logprobs, etc.) to Pydantic AI. | pydantic/ | 21k | — | ~3.2k | Automated safety check: Pass | MIT | today |
| 111 | Prepares many-file, single-turn transforms such as translating or rewriting as a plan, then submits it to the asynchronous, half-price DashScope Batch API through the qwen batch CLI. | QwenLM/ | 28k | — | ~2.2k | Automated safety check: Pass | Apache-2.0 | today |
| 112 | 112.Prompt Caching Cache the parts of the prompt that don't change so a long-running loop stops paying full price on every turn. | Archive228/ | 755 | — | ~735 | Automated safety check: Pass | MIT | 2 mo ago |
| 113 | 113.Cost Tracking A skill your agent uses when the user wants to capture AI tool costs, review spending trends, set cost budgets, or integrate cost data into health snapshots — guides quarterly cost capture, records… | Habitat-Thinking/ | 114 | — | ~1.7k | Automated safety check: Pass | Unknown | 20 days ago |
| 114 | Specifies how quota-axi reads LLM subscription quota windows, derives pace and runway, computes a selection signal and renders output in TOON, JSON and TUI tiers. | kunchenguid/ | 147 | — | ~1.2k | Automated safety check: Pass | MIT | yesterday |
| 115 | 115.Prompt Caching Caching strategies for LLM prompts including Anthropic prompt caching, response caching, and CAG (Cache Augmented Generation) Use when: prompt caching, cache prompt, response cache, cag, cache… | davila7/ | 33k | 5 repos | ~452 | Automated safety check: Pass | MIT | today |
| 116 | 116.Debug Ohmycode Guide for debugging OhMyCode issues. An agent skill from AlphaLab-USTC/OhMyCode. | AlphaLab-USTC/ | 131 | — | ~1.3k | Automated safety check: Pass | MIT | 6 mo ago |
| 117 | Checks every subagent prompt before spawning, swapping pasted files and context for paths and short summaries and trimming the brief, to avoid multiplied token cost. | LichAmnesia/ | 234 | — | ~1.6k | Automated safety check: Pass | MIT | 4 mo ago |
| 118 | 118.Cost Tracker 追蹤並計算當前 session 的 Token 使用量和花費(USD)。產出前後對比報告,幫助使用者了解 workspace 優化的實際效益。 | zeuikli/ | 156 | — | ~830 | Automated safety check: Pass | No licence | 5 days ago |
| 119 | 119.Cost Tracking Track and report Claude Code token usage, spending, and budgets from the local ECC cost-tracker metrics log. | affaan-m/ | 277k | 1 repo | ~1.3k | Automated safety check: Pass | MIT | today |
| 120 | Continue's prompt caching is opt-in via config and off by default. | OnlyTerp/ | 114 | — | ~977 | Automated safety check: Pass | Unknown | 1 mo ago |
| 121 | 121.Cost Tracker Track session costs, set budget alerts, and optimize token spend. | rohitg00/ | 2.9k | — | ~814 | Automated safety check: Pass | No licence | 12 days ago |
| 122 | A skill your agent uses when the user asks to optimize prompts, design prompt templates, evaluate LLM outputs with an eval set, measure RAG retrieval quality, validate agent/tool configurations… | alirezarezvani/ | 28k | 1 repo | ~2.5k | Automated safety check: Pass | MIT | 1 mo ago |
| 123 | 123.Morning Briefing Generate a concise daily infrastructure briefing. An agent skill from bolivian-peru/os-moda. | bolivian-peru/ | 119 | — | ~897 | Automated safety check: Pass | Apache-2.0 | 3 mo ago |
| 124 | A skill your agent uses to classify the documentation impact of a pull request diff, returning one of three verdicts -- no-change, in-place edit, or structural change -- with bounded LLM cost. | microsoft/ | 4k | — | ~1.8k | Automated safety check: Pass | MIT | yesterday |
| 125 | This skill should be used when writing, enhancing, or evaluating the launch prompt for a long-running autonomous agent or a parallel multi-agent orchestration attacking a hard problem: pseudo-formal… | guanyang/ | 977 | 1 repo | ~6.4k | Automated safety check: Pass | MIT | today |
| 126 | This skill should be used when a model gets read-write control over its own live context window instead of a harness-scheduled compaction policy: the context exposed as an editable file the model… | guanyang/ | 977 | 1 repo | ~6k | Automated safety check: Pass | MIT | today |
| 127 | 127.Cost AI Build Cost Tracker — track how much AI is costing you per feature. | Houseofmvps/ | 123 | — | ~1.1k | Automated safety check: Notes | MIT | 3 mo ago |
| 128 | 128.Notepad System A skill your agent uses when context compression is imminent, when resuming a session, or when preserving critical decisions across long tasks. | vibeeval/ | 532 | — | ~1.7k | Automated safety check: Pass | MIT | 2 mo ago |
| 129 | 129.Compact Guide Context window management and token optimization guide. An agent skill from jh941213/my-cc-harness. | jh941213/ | 125 | — | ~621 | Automated safety check: Pass | No licence | 2 mo ago |
| 130 | Analyze markdown files for token efficiency and reduce context-window bloat. | Azure/ | 134 | — | ~602 | Automated safety check: Pass | MIT | yesterday |
| 131 | Get RED metrics + service maps + frontend RUM + AI/LLM monitoring out of Grafana Cloud — Application Observability (tracesspanmetrics from OTel traces, p50/p95/p99 latency, exemplar-to-trace… | grafana/ | 282 | — | ~1.8k | Automated safety check: Pass | Apache-2.0 | 2 days ago |
| 132 | Analyze and reduce token consumption in agentic workflows — guardrail-specific entry points, measurement, and optimization techniques. | github/ | 5.4k | — | ~1.2k | Automated safety check: Pass | MIT | today |
| 133 | Monitor LLMs and agentic apps: performance, token/cost, response quality, and workflow orchestration. | aspectrr/ | 405 | — | ~858 | Automated safety check: Pass | MIT | 5 mo ago |
| 134 | Create and manage webhook subscriptions for event-driven agent activation, or for direct push notifications (zero LLM cost). | Tommy-yw/ | 546 | 1 repo | ~1.7k | Automated safety check: Notes | MIT | 4 mo ago |
| 135 | 135.Agentflow Orchestrate autonomous AI development pipelines through your Kanban board (Asana, GitHub Projects, Linear). | sickn33/ | 47k | 2 repos | ~2k | Automated safety check: Pass | MIT | 2 days ago |
| 136 | 136.LLM Caching Implement multi-layer LLM caching with exact match, semantic similarity, and provider-side prompt caching. | sickn33/ | 47k | 2 repos | ~2.8k | Automated safety check: Pass | MIT | 2 days ago |
| 137 | 137.Prompt Caching Caching strategies for LLM prompts including Anthropic prompt caching, response caching, and CAG (Cache Augmented Generation) | sickn33/ | 47k | 2 repos | ~3.4k | Automated safety check: Pass | MIT | 2 days ago |
| 138 | OpenCode places cachePoint on Bedrock messages containing DocumentBlocks, which produces a "nothing available to cache" error. | OnlyTerp/ | 114 | — | ~966 | Automated safety check: Pass | Unknown | 1 mo ago |
| 139 | Optimize LLM prompt caching hit rate to reduce API costs and improve latency. | affaan-m/ | 276k | — | ~2.5k | Automated safety check: Pass | MIT | yesterday |
| 140 | 140.Repo Map Ranked symbol map of a codebase within a token budget — a compact "what matters in this repo" before reading files. | AnastasiyaW/ | 154 | — | ~1.5k | Automated safety check: Pass | MIT | yesterday |
| 141 | Master the four operations of context engineering — Write, Select, Compress, Isolate. | rohitg00/ | 2.9k | — | ~1.6k | Automated safety check: Pass | No licence | 12 days ago |
| 142 | Send images, audio, video, or documents into an AG2 beta Agent alongside text. | ag2ai/ | 252 | — | ~1.7k | Automated safety check: Pass | Apache-2.0 | 1 mo ago |
| 143 | Investigate LLM spend in PostHog — total cost over time, cost by model, provider, user, trace, or custom dimension, token and cache-hit economics, and cost regressions. | PostHog/ | 40k | — | ~2.9k | Automated safety check: Pass | Unknown | yesterday |
| 144 | 144.Hivemind Orchestrate free opencode workers from Claude Code to cut token costs. | alirezarezvani/ | 28k | — | ~2.4k | Automated safety check: Pass | Apache-2.0 | 1 mo ago |