Topic · AI & LLM Engineering
Best LLM cost and token optimization skills, page 5
LLM cost and token optimization skills, ranked
Ranked by score. Sort bymost stars,trending,newest,recently updated
| # | Skill | Repository | Stars | Used in | Tokens | Auto-check | Licence | Updated |
|---|---|---|---|---|---|---|---|---|
| 193 | Analyze markdown files for token efficiency and reduce context-window bloat. | Azure/ | 134 | — | ~494 | Automated safety check: Pass | MIT | today |
| 194 | Monitoring and observability patterns for Prometheus metrics, Grafana dashboards, Langfuse v4 LLM tracing (astype, scorecurrentspan, shouldexportspan, LangfuseMedia), and drift detection. | yonatangross/ | 289 | — | ~2.2k | Automated safety check: Pass | MIT | today |
| 195 | 195.LLM Router Selects the optimal LLM model and provider for each task based on complexity, cost budget, and capability requirements. | curiositech/ | 243 | 1 repo | ~1.7k | Automated safety check: Pass | MIT | 1 mo ago |
| 196 | 196.Dt Obs Genai Analyze & debug GenAI/LLM apps: token cost & caching by prompt, model & provider; latency/errors; agent & tool loops/failures; conversations; guardrails; evaluations; OpenTelemetry/dt-evals setup. | Dynatrace/ | 161 | — | ~4.5k | Automated safety check: Pass | Apache-2.0 | 6 days ago |
| 197 | 197.Claude API A skill your agent uses when building or debugging apps that call the Claude API — implementing tool use, streaming, vision, prompt caching, batch processing, extended thinking, or an agentic loop… | kid-sid/ | 189 | — | ~2.7k | Automated safety check: Pass | MIT | 2 mo ago |
| 198 | 198.Token Optimizer Reduce OpenClaw token usage and API costs through smart model routing, heartbeat optimization, budget tracking, and native 2026.2.15 features (session pruning, bootstrap size limits, cache TTL… | LeoYeAI/ | 2.2k | — | ~5.2k | Automated safety check: Pass | MIT | 2 mo ago |
| 199 | 199.Zero Token Zero token cost. An agent skill from LeoYeAI/openclaw-master-skills. | LeoYeAI/ | 2.2k | — | ~889 | Automated safety check: Pass | MIT | 2 mo ago |
| 200 | Replay an LLM inference request trace (Mooncake / vLLM / SGLang hashids format) against a block-level KV prefix cache and compute hit statistics. | benchflow-ai/ | 1.8k | — | ~2.3k | Automated safety check: Pass | Apache-2.0 | 2 mo ago |
| 201 | Track real-time API cost accrual during LLM execution. An agent skill from curiositech/some_claude_skills. | curiositech/ | 243 | 1 repo | ~2.3k | Automated safety check: Pass | MIT | 1 mo ago |
| 202 | Audit LLM token cost estimates against actual API usage. An agent skill from curiositech/some_claude_skills. | curiositech/ | 243 | 1 repo | ~1.4k | Automated safety check: Notes | MIT | 1 mo ago |
| 203 | 203.RAG Eval Iterate on RAG systems with structured evals instead of eyeballing. | glebis/ | 389 | — | ~1.5k | Automated safety check: Pass | MIT | 11 days ago |
| 204 | Anthropic API prompt caching: TTL, breakpoints, stacking, invalidation, hit rate. | softspark/ | 179 | — | ~1.1k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 205 | 205.Token Cost Measure before optimizing — estimate token counts locally with stated heuristics, price them at your model's rates, and quantify before/after savings, because token optimization without measurement… | mohitagw15856/ | 1.4k | — | ~1.3k | Automated safety check: Pass | MIT | today |
| 206 | Analyzes markdown files for token efficiency. An agent skill from microsoft/GitHub-Copilot-for-Azure. | microsoft/ | 255 | — | ~364 | Automated safety check: Pass | MIT | today |
| 207 | Optimize Anthropic API costs — model selection, prompt caching, batches, Use when working with cost-tuning patterns. | jeremylongshore/ | 2.8k | — | ~1.3k | Automated safety check: Pass | MIT | today |
| 208 | Optimize Anthropic API latency — streaming, prompt caching, model selection, Use when working with performance-tuning patterns. | jeremylongshore/ | 2.8k | — | ~1.3k | Automated safety check: Pass | MIT | today |
| 209 | Analyze and control Flexport request volume without inventing undocumented global limits or headers. | jeremylongshore/ | 2.8k | — | ~1.1k | Automated safety check: Pass | MIT | today |
| 210 | A skill your agent uses when you need to keep PII out of Groq API calls, filter model responses, audit-log conversations, or track token cost and usage for a Groq integration. | jeremylongshore/ | 2.8k | — | ~1.3k | Automated safety check: Notes | MIT | today |
| 211 | Set up observability for Groq integrations: latency histograms, token throughput, rate limit gauges, cost tracking, and Prometheus alerts. | jeremylongshore/ | 2.8k | — | ~1.6k | Automated safety check: Pass | MIT | today |
| 212 | Optimize context window usage for OpenRouter models to reduce cost and improve quality. | jeremylongshore/ | 2.8k | — | ~2.4k | Automated safety check: Pass | MIT | today |
| 213 | 213.AI Observability Implement comprehensive observability for LLM applications including tracing (Langfuse/Helicone), cost tracking, token optimization, RAG evaluation metrics (RAGAS), hallucination detection, and… | omer-metin/ | 162 | — | ~578 | Automated safety check: Pass | Apache-2.0 | 8 mo ago |
| 214 | 214.Context Engine Context management engine for AI coding agents. An agent skill from borghei/Claude-Skills. | borghei/ | 881 | — | ~2.2k | Automated safety check: Pass | MIT | today |
| 215 | 215.Pp Posthog Every PostHog resource in one CLI — with offline search, agent-native output, and cross-resource analytics no... | mvanhorn/ | 2.1k | — | ~3.2k | Automated safety check: Notes | Apache-2.0 | today |
| 216 | Track Clawdbot AI model usage and estimate costs. An agent skill from sundial-org/awesome-openclaw-skills. | sundial-org/ | 663 | — | ~622 | Automated safety check: Pass | No licence | 7 mo ago |
| 217 | Preserve conversation continuity across token compaction cycles by extracting and archiving all prompts with date-wise entries. | sundial-org/ | 663 | — | ~906 | Automated safety check: Pass | No licence | 7 mo ago |
| 218 | 218.Cco Pack Build an optimal context pack for the user's task — ranked file list with offset/limit suggestions, based on git state, mentioned paths, and historical patterns | egorfedorov/ | 114 | — | ~448 | Automated safety check: Pass | MIT | 8 days ago |
| 219 | Review skill PRs with structured severity-rated feedback covering token budgets, routing conflicts, and repo conventions. | microsoft/ | 255 | — | ~482 | Automated safety check: Pass | MIT | today |
| 220 | 220.Cost Tracking A skill your agent uses when metering and capping AI or cloud app spend — tokens read from the response usage object, priced off a dated rate table, ledgered per user/tenant/feature, with alerts and… | ericrisco/ | 167 | — | ~3.2k | Automated safety check: Pass | MIT | today |
| 221 | A skill your agent uses when calling open-weight LLMs on Together AI or Fireworks AI's OpenAI-compatible endpoints — baseurl plus namespaced model id, the cheapest model that clears the bar… | ericrisco/ | 167 | — | ~3.3k | Automated safety check: Pass | MIT | today |
| 222 | Real-time cost tracking, budget enforcement, and ROI measurement for AI agent operations. | majiayu000/ | 666 | 1 repo | ~5k | Automated safety check: Notes | MIT | today |
| 223 | Comprehensive cost tracking and optimization for production Claude deployments. | majiayu000/ | 666 | 1 repo | ~3.1k | Automated safety check: Pass | MIT | today |
| 224 | Compress selected context to a target token budget while preserving decisions, evidence, constraints, and unresolved questions. | majiayu000/ | 666 | 1 repo | ~2.2k | Automated safety check: Pass | MIT | today |
| 225 | Compact one or more Spec-Driven Development artifacts in place to reduce token cost on subsequent /speckit. | haru/ | 106 | — | ~1.8k | Automated safety check: Pass | MIT | today |
| 226 | Toggle a project-local concise-output directive that suppresses agent prose padding during SDD steps. | haru/ | 106 | — | ~1.9k | Automated safety check: Pass | MIT | today |
| 227 | Build a focused reading manifest for the next workflow step. | haru/ | 106 | — | ~1.3k | Automated safety check: Pass | MIT | today |
| 228 | Show token usage for every SDD artifact in the active feature, the projected context window for each upcoming phase, and the savings already realized by token-budget (full backups vs current files). | haru/ | 106 | — | ~1k | Automated safety check: Pass | MIT | today |
| 229 | LLM API 使用成本优化模式——基于任务复杂度的模型路由、预算跟踪、重试逻辑和提示词缓存. An agent skill from xu-xiang/everything-claude-code-zh. | xu-xiang/ | 2k | — | ~1.1k | Automated safety check: Pass | MIT | 7 mo ago |
| 230 | Clay workflows — reduce a workflow's credit and LLM cost via the CLI (clay workflows commands). | clay-run/ | 130 | — | ~1.3k | Automated safety check: Pass | No licence | today |
| 231 | Turn JSON or PostgreSQL jsonb payloads into compact readable context for LLMs. | aiskillstore/ | 430 | — | ~1k | Automated safety check: Pass | No licence | today |
| 232 | Audit and optimize OpenClaw token usage, cron job efficiency, and agent performance. | LeoYeAI/ | 2.2k | — | ~4.5k | Automated safety check: Notes | MIT | 2 mo ago |
| 233 | 233.Anth Cost Tuning Optimize Anthropic Claude API costs with model routing, prompt caching, batching, and spend monitoring. | jeremylongshore/ | 2.8k | — | ~2.1k | Automated safety check: Pass | MIT | today |
| 234 | Optimize Claude API performance with prompt caching, model selection, streaming, and latency reduction techniques. | jeremylongshore/ | 2.8k | — | ~1.9k | Automated safety check: Pass | MIT | today |
| 235 | 235.Venice Chat Call POST /chat/completions on Venice. An agent skill from veniceai/skills. | veniceai/ | 143 | — | ~5.9k | Automated safety check: Pass | MIT | 2 days ago |
| 236 | Compresses artifacts for judge evaluation. An agent skill from closedloop-ai/claude-plugins. | closedloop-ai/ | 122 | — | ~2.1k | Automated safety check: Notes | Apache-2.0 | today |
| 237 | 237.Cost Estimation A skill your agent uses when the user wants to estimate or predict the cost, token usage, or time of a task BEFORE it runs — "how much will this feature cost to build", "estimate the tokens for this… | Habitat-Thinking/ | 114 | — | ~5.4k | Automated safety check: Pass | Unknown | 17 days ago |
| 238 | 238.Pp Groq Every Groq endpoint in your terminal, plus a local ledger that tracks token cost and rate-limit budget. | mvanhorn/ | 2.1k | — | ~7.9k | Automated safety check: Notes | Apache-2.0 | today |
| 239 | Review what an LLM feature or agent actually puts in its context window — and find what's bloating, missing, or fighting itself. | mohitagw15856/ | 1.4k | — | ~1.4k | Automated safety check: Pass | MIT | today |
| 240 | Model the cost and latency of an LLM feature before it ships and surprises the bill. | mohitagw15856/ | 1.4k | — | ~985 | Automated safety check: Pass | MIT | today |
Explore related skills
Category
More topics in AI & LLM Engineering
- Building AI agents563
- Deep learning415
- Embeddings386
- LLM inference and serving372
- Prompt engineering360
- Retrieval-augmented generation358
- Fine-tuning309
- LLM evaluation308
- Speech recognition and synthesis308
- Structured output and tool calling276
- LLM API integration255
- Model routing and gateways255
- LLM observability240
- LLM guardrails221
- Computer vision203
- Model hubs and datasets180
- GPU and accelerator computing176
- Diffusion and image models166
- Natural language processing131
- Reinforcement learning66
- AI interpretability23