Search
LLM cost and token optimization
Skills
Sort:BestMost starsTrending todayTrending this weekTrending this monthNewestRecently updatedName
| # | Skill | Repository | Stars | Used in | Tokens | Auto-check | Licence | Updated |
|---|---|---|---|---|---|---|---|---|
| 193 | 193.Dt Obs Genai Analyze & debug GenAI/LLM apps: token cost & caching by prompt, model & provider; latency/errors; agent & tool loops/failures; conversations; guardrails; evaluations; OpenTelemetry/dt-evals setup. | Dynatrace/ | 163 | — | ~4.5k | Automated safety check: Pass | Apache-2.0 | 10 days ago |
| 194 | 194.Claude API A skill your agent uses when building or debugging apps that call the Claude API — implementing tool use, streaming, vision, prompt caching, batch processing, extended thinking, or an agentic loop… | kid-sid/ | 190 | — | ~2.7k | Automated safety check: Pass | MIT | 2 mo ago |
| 195 | 195.Token Optimizer Reduce OpenClaw token usage and API costs through smart model routing, heartbeat optimization, budget tracking, and native 2026.2.15 features (session pruning, bootstrap size limits, cache TTL… | LeoYeAI/ | 2.2k | — | ~5.2k | Automated safety check: Pass | MIT | 2 mo ago |
| 196 | 196.Zero Token Zero token cost. An agent skill from LeoYeAI/openclaw-master-skills. | LeoYeAI/ | 2.2k | — | ~889 | Automated safety check: Pass | MIT | 2 mo ago |
| 197 | A skill your agent uses for debugging DSPy programs, inspecthistory, tracing LLM calls, custom callbacks, observability, monitoring, and cost tracking. | OmidZamani/ | 123 | — | ~2.1k | Automated safety check: Warn | MIT | 3 mo ago |
| 198 | Replay an LLM inference request trace (Mooncake / vLLM / SGLang hashids format) against a block-level KV prefix cache and compute hit statistics. | benchflow-ai/ | 1.8k | — | ~2.3k | Automated safety check: Pass | Apache-2.0 | 2 mo ago |
| 199 | 199.RAG Eval Iterate on RAG systems with structured evals instead of eyeballing. | glebis/ | 391 | — | ~1.5k | Automated safety check: Pass | MIT | 3 days ago |
| 200 | Anthropic API prompt caching: TTL, breakpoints, stacking, invalidation, hit rate. | softspark/ | 179 | — | ~1.1k | Automated safety check: Pass | Apache-2.0 | 3 days ago |
| 201 | 201.Token Cost Measure before optimizing — estimate token counts locally with stated heuristics, price them at your model's rates, and quantify before/after savings, because token optimization without measurement… | mohitagw15856/ | 1.4k | — | ~1.3k | Automated safety check: Pass | MIT | 2 days ago |
| 202 | Analyzes markdown files for token efficiency. An agent skill from microsoft/GitHub-Copilot-for-Azure. | microsoft/ | 255 | — | ~364 | Automated safety check: Pass | MIT | yesterday |
| 203 | Monitor and optimize LLM costs using Langfuse analytics and dashboards. | jeremylongshore/ | 2.8k | — | ~2.4k | Automated safety check: Pass | MIT | yesterday |
| 204 | Optimize Anthropic API costs — model selection, prompt caching, batches, Use when working with cost-tuning patterns. | jeremylongshore/ | 2.8k | — | ~1.3k | Automated safety check: Pass | MIT | yesterday |
| 205 | Optimize Anthropic API latency — streaming, prompt caching, model selection, Use when working with performance-tuning patterns. | jeremylongshore/ | 2.8k | — | ~1.3k | Automated safety check: Pass | MIT | yesterday |
| 206 | Analyze and control Flexport request volume without inventing undocumented global limits or headers. | jeremylongshore/ | 2.8k | — | ~1.1k | Automated safety check: Pass | MIT | yesterday |
| 207 | A skill your agent uses when you need to keep PII out of Groq API calls, filter model responses, audit-log conversations, or track token cost and usage for a Groq integration. | jeremylongshore/ | 2.8k | — | ~1.3k | Automated safety check: Notes | MIT | yesterday |
| 208 | Set up observability for Groq integrations: latency histograms, token throughput, rate limit gauges, cost tracking, and Prometheus alerts. | jeremylongshore/ | 2.8k | — | ~1.6k | Automated safety check: Pass | MIT | yesterday |
| 209 | Optimize context window usage for OpenRouter models to reduce cost and improve quality. | jeremylongshore/ | 2.8k | — | ~2.4k | Automated safety check: Pass | MIT | yesterday |
| 210 | 210.AI Observability Implement comprehensive observability for LLM applications including tracing (Langfuse/Helicone), cost tracking, token optimization, RAG evaluation metrics (RAGAS), hallucination detection, and… | omer-metin/ | 163 | — | ~578 | Automated safety check: Pass | Apache-2.0 | 8 mo ago |
| 211 | 211.Context Engine Context management engine for AI coding agents. An agent skill from borghei/Claude-Skills. | borghei/ | 891 | — | ~2.2k | Automated safety check: Pass | MIT | 4 days ago |
| 212 | 212.Pp Posthog Every PostHog resource in one CLI — with offline search, agent-native output, and cross-resource analytics no... | mvanhorn/ | 2.1k | — | ~3.2k | Automated safety check: Notes | Apache-2.0 | yesterday |
| 213 | 213.Cco Pack Build an optimal context pack for the user's task — ranked file list with offset/limit suggestions, based on git state, mentioned paths, and historical patterns | egorfedorov/ | 114 | — | ~448 | Automated safety check: Pass | MIT | 11 days ago |
| 214 | 214.Cost Tracking A skill your agent uses when metering and capping AI or cloud app spend — tokens read from the response usage object, priced off a dated rate table, ledgered per user/tenant/feature, with alerts and… | ericrisco/ | 180 | — | ~3.2k | Automated safety check: Pass | MIT | yesterday |
| 215 | A skill your agent uses when calling open-weight LLMs on Together AI or Fireworks AI's OpenAI-compatible endpoints — baseurl plus namespaced model id, the cheapest model that clears the bar… | ericrisco/ | 180 | — | ~3.3k | Automated safety check: Pass | MIT | yesterday |
| 216 | Track Clawdbot AI model usage and estimate costs. An agent skill from sundial-org/awesome-openclaw-skills. | sundial-org/ | 663 | — | ~622 | Automated safety check: Pass | No licence | 7 mo ago |
| 217 | Preserve conversation continuity across token compaction cycles by extracting and archiving all prompts with date-wise entries. | sundial-org/ | 663 | — | ~906 | Automated safety check: Pass | No licence | 7 mo ago |
| 218 | Review skill PRs with structured severity-rated feedback covering token budgets, routing conflicts, and repo conventions. | microsoft/ | 255 | — | ~482 | Automated safety check: Pass | MIT | yesterday |
| 219 | Compact one or more Spec-Driven Development artifacts in place to reduce token cost on subsequent /speckit. | haru/ | 106 | — | ~1.8k | Automated safety check: Pass | MIT | yesterday |
| 220 | Toggle a project-local concise-output directive that suppresses agent prose padding during SDD steps. | haru/ | 106 | — | ~1.9k | Automated safety check: Pass | MIT | yesterday |
| 221 | Build a focused reading manifest for the next workflow step. | haru/ | 106 | — | ~1.3k | Automated safety check: Pass | MIT | yesterday |
| 222 | Show token usage for every SDD artifact in the active feature, the projected context window for each upcoming phase, and the savings already realized by token-budget (full backups vs current files). | haru/ | 106 | — | ~1k | Automated safety check: Pass | MIT | yesterday |
| 223 | LLM API 使用成本优化模式——基于任务复杂度的模型路由、预算跟踪、重试逻辑和提示词缓存. An agent skill from xu-xiang/everything-claude-code-zh. | xu-xiang/ | 2k | — | ~1.1k | Automated safety check: Pass | MIT | 7 mo ago |
| 224 | Clay workflows — reduce a workflow's credit and LLM cost via the CLI (clay workflows commands). | clay-run/ | 132 | — | ~1.3k | Automated safety check: Pass | No licence | 3 days ago |
| 225 | Turn JSON or PostgreSQL jsonb payloads into compact readable context for LLMs. | aiskillstore/ | 433 | — | ~1k | Automated safety check: Pass | No licence | yesterday |
| 226 | Audit and optimize OpenClaw token usage, cron job efficiency, and agent performance. | LeoYeAI/ | 2.2k | — | ~4.5k | Automated safety check: Notes | MIT | 2 mo ago |
| 227 | 227.Anth Cost Tuning Optimize Anthropic Claude API costs with model routing, prompt caching, batching, and spend monitoring. | jeremylongshore/ | 2.8k | — | ~2.1k | Automated safety check: Pass | MIT | yesterday |
| 228 | Optimize Claude API performance with prompt caching, model selection, streaming, and latency reduction techniques. | jeremylongshore/ | 2.8k | — | ~1.9k | Automated safety check: Pass | MIT | yesterday |
| 229 | Reduce LLM API and infrastructure costs through model selection, prompt caching, batching, caching, quantization, and self-hosting strategies. | BagelHole/ | 1.2k | — | ~2.2k | Automated safety check: Pass | MIT | 4 mo ago |
| 230 | 230.Venice Chat Call POST /chat/completions on Venice. An agent skill from veniceai/skills. | veniceai/ | 144 | — | ~5.9k | Automated safety check: Pass | MIT | 5 days ago |
| 231 | Compresses artifacts for judge evaluation. An agent skill from closedloop-ai/claude-plugins. | closedloop-ai/ | 122 | — | ~2.1k | Automated safety check: Notes | Apache-2.0 | yesterday |
| 232 | 232.Cost Estimation A skill your agent uses when the user wants to estimate or predict the cost, token usage, or time of a task BEFORE it runs — "how much will this feature cost to build", "estimate the tokens for this… | Habitat-Thinking/ | 114 | — | ~5.4k | Automated safety check: Pass | Unknown | 20 days ago |
| 233 | 233.Pp Groq Every Groq endpoint in your terminal, plus a local ledger that tracks token cost and rate-limit budget. | mvanhorn/ | 2.1k | — | ~7.9k | Automated safety check: Notes | Apache-2.0 | yesterday |
| 234 | Review what an LLM feature or agent actually puts in its context window — and find what's bloating, missing, or fighting itself. | mohitagw15856/ | 1.4k | — | ~1.4k | Automated safety check: Pass | MIT | 2 days ago |
| 235 | Model the cost and latency of an LLM feature before it ships and surprises the bill. | mohitagw15856/ | 1.4k | — | ~985 | Automated safety check: Pass | MIT | 2 days ago |
| 236 | Choose the right LLM for a task by trading off quality, cost, latency, and constraints. | mohitagw15856/ | 1.4k | — | ~1.1k | Automated safety check: Pass | MIT | 2 days ago |
| 237 | 237.Token Diet Cut LLM output tokens 40–70% by stripping grammatical scaffolding while preserving every fact — telegraphic output modes, when they pay (pipelines, long sessions) and when they don't (single shots… | mohitagw15856/ | 1.4k | — | ~1.5k | Automated safety check: Pass | MIT | 2 days ago |
| 238 | 238.Cost Optimizer Ultimate cost optimization toolkit for OpenClaw/Claude Code. | LeoYeAI/ | 2.2k | — | ~3.3k | Automated safety check: Notes | MIT | 2 mo ago |
| 239 | Java/Maven single-module deep documentation generator. An agent skill from LeoYeAI/openclaw-master-skills. | LeoYeAI/ | 2.2k | — | ~3.4k | Automated safety check: Pass | MIT | 2 mo ago |
| 240 | 240.SDK Core A skill your agent uses for direct LiteLLM Python SDK work: chat/text completions, async calls, streaming, embeddings, structured outputs, tools, token/cost checks, caching, callbacks, import/smoke… | VectorSpaceLab/ | 331 | — | ~1.1k | Automated safety check: Pass | Unknown | 1 mo ago |