Search
LLM cost and token optimization
Skills
Sort:BestMost starsTrending todayTrending this weekTrending this monthNewestRecently updatedName
| # | Skill | Repository | Stars | Used in | Tokens | Auto-check | Licence | Updated |
|---|---|---|---|---|---|---|---|---|
| 145 | Configure and optimize .geminiignore files for AI context window efficiency and token cost reduction (FinOps). | sickn33/ | 47k | 1 repo | ~1.4k | Automated safety check: Notes | MIT | 2 days ago |
| 146 | Reduce LLM API and infrastructure costs through model selection, prompt caching, batching, caching, quantization, and self-hosting strategies. | sickn33/ | 47k | 1 repo | ~2.5k | Automated safety check: Pass | MIT | 2 days ago |
| 147 | 147.LLM Gateway Deploy an API gateway for LLM traffic with load balancing, rate limiting, key management, semantic caching, fallback routing, and cost tracking. | sickn33/ | 47k | 1 repo | ~2.1k | Automated safety check: Pass | MIT | 2 days ago |
| 148 | Optimizes AI agent performance by pruning redundant context, managing token usage, and enforcing ultra-concise, direct-to-value responses. | sickn33/ | 47k | 1 repo | ~1.2k | Automated safety check: Pass | MIT | 2 days ago |
| 149 | 149.Zipai Optimizer Ultra-dense token optimizer skill for prompt caching, log pruning, AST-based inspection, and minified JSON payloads. | sickn33/ | 47k | 1 repo | ~1.3k | Automated safety check: Pass | MIT | 2 days ago |
| 150 | 150.Cost Tracking ローカルのコスト追跡データベースからClaude Codeのトークン使用量、支出、予算を追跡・レポートします。コスト、支出、使用量、トークン、予算、またはプロジェクト、ツール、セッション、日付によるコスト内訳について質問する場合に使用します。 | affaan-m/ | 276k | — | ~829 | Automated safety check: Pass | MIT | yesterday |
| 151 | 回答する前に、どれだけの回答深度を消費するかについてユーザーに情報に基づいた選択を提供する。ユーザーが回答の長さ、深さ、またはトークンバジェットを明示的に制御したい場合にこのスキルを使用する。トリガー条件:"token budget", "token count", "token usage", "token limit", "response length", "answer depth"… | affaan-m/ | 276k | — | ~910 | Automated safety check: Pass | MIT | yesterday |
| 152 | 在回答前,为用户提供关于消耗多少响应深度的知情选择。当用户明确希望控制响应长度、深度或令牌预算时使用此技能。触发条件:"token budget", "token count", "token usage", "token limit", "response length", "answer depth", "short version", "brief answer", "detailed… | affaan-m/ | 276k | — | ~927 | Automated safety check: Pass | MIT | yesterday |
| 153 | A skill your agent uses when the user asks to "optimize CLAUDE.md", "create a new skill", "write a custom agent", "configure hooks", "manage context window", "set up MCP servers", "scaffold a skill… | borghei/ | 891 | — | ~1.9k | Automated safety check: Pass | MIT | 4 days ago |
| 154 | 154.Token Efficiency Reduce token waste by 40-60% through anti-sycophancy rules, tool-call budgets, one-pass coding, task profiles, and read-before-write enforcement. | rohitg00/ | 2.9k | — | ~1k | Automated safety check: Pass | No licence | 12 days ago |
| 155 | 155.Workflow Mastery Claude Code workflow mastery for .NET developers. An agent skill from codewithmukesh/dotnet-claude-kit. | codewithmukesh/ | 756 | — | ~3.5k | Automated safety check: Pass | MIT | 2 mo ago |
| 156 | Analyze the most expensive users in AI observability and explain why they cost so much. | PostHog/ | 40k | — | ~3.9k | Automated safety check: Pass | Unknown | yesterday |
| 157 | 157.RAG Architect A skill your agent uses when the user asks to design a RAG pipeline, choose a chunking strategy or embedding model, pick a vector database, or evaluate retrieval quality (precision@k, recall@k, NDCG). | alirezarezvani/ | 28k | — | ~1.1k | Automated safety check: Pass | MIT | 1 mo ago |
| 158 | OpenCode doesn't set promptcachekey for Mistral models, missing the documented 10% cached-token discount. | OnlyTerp/ | 114 | — | ~638 | Automated safety check: Pass | Unknown | 1 mo ago |
| 159 | Read accumulated cost-tracking spend + budget config, compute utilization, emit 50/75/90/100% alert ladder | ruvnet/ | 74k | — | ~645 | Automated safety check: Notes | MIT | yesterday |
| 160 | 160.Cost Burn Burn-rate trend over time with optional drift-alert exit code. | ruvnet/ | 74k | — | ~824 | Automated safety check: Notes | MIT | yesterday |
| 161 | 161.Cost Export Export cost-tracking telemetry in Prometheus textfile or webhook JSON formats — for external observability (Grafana, Datadog, custom dashboards) | ruvnet/ | 74k | — | ~687 | Automated safety check: Notes | MIT | yesterday |
| 162 | 162.Cost Optimize Analyze token usage patterns and recommend cost optimizations with estimated savings | ruvnet/ | 74k | — | ~997 | Automated safety check: Notes | MIT | yesterday |
| 163 | 163.Cost Report Generate a cost report showing token usage and USD costs by agent and model | ruvnet/ | 74k | — | ~830 | Automated safety check: Notes | MIT | yesterday |
| 164 | 164.Cost Summary Single-shot programmatic dump of all cost data — total spend, per-tier, top session, budget status, federation aggregate. | ruvnet/ | 74k | — | ~594 | Automated safety check: Notes | MIT | yesterday |
| 165 | 165.Cost Track Auto-capture per-session token usage from the Claude Code session jsonl and persist to the cost-tracking namespace | ruvnet/ | 74k | — | ~773 | Automated safety check: Notes | MIT | yesterday |
| 166 | Intercept the response flow to offer the user a choice about response depth before Claude answers Offers the user an informed choice about how much response depth to consume before answering. | aAAaqwq/ | 105 | 4 repos | ~1.5k | Automated safety check: Pass | MIT | 2 days ago |
| 167 | Analyze and reduce LLM spend: read usage breakdowns by call site, model, and inference profile, understand single-winner profile resolution, and pin call sites to managed profiles (Balanced /… | vellum-ai/ | 1.4k | — | ~3.9k | Automated safety check: Pass | MIT | 2 days ago |
| 168 | 168.Map Structural codebase index generator. An agent skill from SethGammon/Citadel. | SethGammon/ | 924 | — | ~1.8k | Automated safety check: Pass | MIT | yesterday |
| 169 | When agent sessions generate millions of tokens of conversation history, compression becomes mandatory. | aiskillstore/ | 433 | 5 repos | ~3.1k | Automated safety check: Pass | No licence | yesterday |
| 170 | Calculate and display the cost of an Output SDK workflow execution run. | growthxai/ | 442 | — | ~1.4k | Automated safety check: Notes | Apache-2.0 | yesterday |
| 171 | A skill your agent uses when managing prompts in production at scale: versioning prompts, running A/B tests on prompts, building prompt registries, preventing prompt regressions, or creating eval… | alirezarezvani/ | 28k | — | ~2.8k | Automated safety check: Pass | MIT | 1 mo ago |
| 172 | 172.Cv Usage A skill your agent uses when the user asks about usage analytics, statistics, token usage, or cost summary — e.g. | tombelieber/ | 111 | — | ~3.1k | Automated safety check: Pass | MIT | 10 days ago |
| 173 | Per-conversation cost view — list every session in cost-tracking with started-at, message count, top model, and total cost | ruvnet/ | 74k | — | ~407 | Automated safety check: Notes | MIT | yesterday |
| 174 | Cost optimization patterns for LLM API usage — model routing by task complexity, budget tracking, retry logic, and prompt caching. | aAAaqwq/ | 105 | 5 repos | ~1.4k | Automated safety check: Pass | MIT | 2 days ago |
| 175 | A skill your agent uses when the user says 'token optimization', 'save tokens', 'context window', 'reduce tokens', 'token stack', or 'TokenStack', or asks about extending context window capacity. | cwinvestments/ | 423 | — | ~1.4k | Automated safety check: Pass | Proprietary | 14 days ago |
| 176 | [omh] Context window or token budget at risk: plan compact context, token/cost budgets, summarization checkpoints, and overflow recovery before long agent work. | rlaope/ | 3.2k | — | ~2k | Automated safety check: Pass | MIT | yesterday |
| 177 | 177.Token Optimizer Reduce OpenClaw token usage and API costs through smart model routing, heartbeat optimization, budget tracking, and multi-provider fallbacks. | LeoYeAI/ | 2.2k | — | ~4.2k | Automated safety check: Pass | MIT | 2 mo ago |
| 178 | Execute this skill optimizes prompts for large language models (llms) to reduce token usage, lower costs, and improve performance. | jeremylongshore/ | 2.8k | — | ~1k | Automated safety check: Pass | MIT | yesterday |
| 179 | Apply Sun Tzu's Art of War to AI agent orchestration. An agent skill from LeoYeAI/openclaw-master-skills. | LeoYeAI/ | 2.2k | — | ~5k | Automated safety check: Pass | MIT | 2 mo ago |
| 180 | Builds generative AI applications on Amazon Bedrock. An agent skill from aws/agent-toolkit-for-aws. | aws/ | 2.8k | — | ~8.6k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 181 | This skill should be used when the user asks to "estimate LLM costs", "count tokens in prompts", "optimize prompt token usage", "compare model pricing", or "reduce LLM API costs". | borghei/ | 891 | — | ~1.1k | Automated safety check: Pass | MIT | 4 days ago |
| 182 | Use proactively whenever LLM API costs come up -- or should. | alirezarezvani/ | 28k | — | ~2.9k | Automated safety check: Pass | MIT | 1 mo ago |
| 183 | 183.Cost Model Standardized cost-estimation framework for greatcto plans. An agent skill from avelikiy/great_cto. | avelikiy/ | 102 | — | ~1.3k | Automated safety check: Pass | MIT | yesterday |
| 184 | Query and analyze a Dynatrace tenant's ACTUAL billing and usage data with DQL against dt.system.events — DPS consumption breakdown, cost-normalized spend ranking, included volume deduction… | Dynatrace/ | 163 | — | ~5.7k | Automated safety check: Pass | Apache-2.0 | 10 days ago |
| 185 | [omh] Tracking cost, tokens, latency, or service health: prepare an operations command-board for wrapper-safe token, cost, latency, run history, queue, failure-mode, external metric-provider, and… | rlaope/ | 3.2k | — | ~2k | Automated safety check: Pass | MIT | yesterday |
| 186 | 186.Ulw Perf [omh] Software slowness, memory leaks, or cost spikes: find where a system is actually slow, leaking, or expensive across runtime, memory, token cost, storage, rendering, inference, CI, and query… | rlaope/ | 3.2k | — | ~2.4k | Automated safety check: Pass | MIT | yesterday |
| 187 | A Senior AI Engineer interviewer that simulates a technical interview focused on prompt engineering and LLM architecture at scale. | PrepLabsAI/ | 112 | — | ~5k | Automated safety check: Pass | MIT | 4 days ago |
| 188 | 188.Oc Doctor Runs a comprehensive 11-section health check on local OpenClaw installations. | LeoYeAI/ | 2.2k | — | ~4.1k | Automated safety check: Pass | MIT | 2 mo ago |
| 189 | Instrument an application with Sentry — detect the platform, install and initialize the SDK if needed, and wire up any signal — error monitoring, tracing/performance, logging, metrics, profiling… | getsentry/ | 268 | — | ~3.2k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 190 | A skill your agent uses when a complex task will span turns or sessions and needs structured working notes to survive context compression or handoff; short single-turn work does not trigger it. | Peiiii/ | 260 | — | ~509 | Automated safety check: Pass | MIT | 2 days ago |
| 191 | Analyze markdown files for token efficiency and reduce context-window bloat. | Azure/ | 134 | — | ~494 | Automated safety check: Pass | MIT | yesterday |
| 192 | Monitoring and observability patterns for Prometheus metrics, Grafana dashboards, Langfuse v4 LLM tracing (astype, scorecurrentspan, shouldexportspan, LangfuseMedia), and drift detection. | yonatangross/ | 292 | — | ~2.2k | Automated safety check: Pass | MIT | yesterday |