Topic · AI & LLM Engineering

Best LLM cost and token optimization skills, page 5

Skills #193–240 of 259, ranked by score.

LLM cost and token optimization skills, ranked

Ranked by score. Sort bymost stars,trending,newest,recently updated

LLM cost and token optimization skills, ranked
#SkillRepositoryStarsUsed inTokensAuto-checkLicenceUpdated
193

Analyze markdown files for token efficiency and reduce context-window bloat.

Azure/azure-sdk-tools134—~494Automated safety check: PassMITtoday
194

Monitoring and observability patterns for Prometheus metrics, Grafana dashboards, Langfuse v4 LLM tracing (astype, scorecurrentspan, shouldexportspan, LangfuseMedia), and drift detection.

yonatangross/orchestkit289—~2.2kAutomated safety check: PassMITtoday
195

Selects the optimal LLM model and provider for each task based on complexity, cost budget, and capability requirements.

curiositech/some_claude_skills2431 repo~1.7kAutomated safety check: PassMIT1 mo ago
196

Analyze & debug GenAI/LLM apps: token cost & caching by prompt, model & provider; latency/errors; agent & tool loops/failures; conversations; guardrails; evaluations; OpenTelemetry/dt-evals setup.

Dynatrace/dynatrace-for-ai161—~4.5kAutomated safety check: PassApache-2.06 days ago
197

A skill your agent uses when building or debugging apps that call the Claude API — implementing tool use, streaming, vision, prompt caching, batch processing, extended thinking, or an agentic loop…

kid-sid/claude-spellbook189—~2.7kAutomated safety check: PassMIT2 mo ago
198

Reduce OpenClaw token usage and API costs through smart model routing, heartbeat optimization, budget tracking, and native 2026.2.15 features (session pruning, bootstrap size limits, cache TTL…

LeoYeAI/openclaw-master-skills2.2k—~5.2kAutomated safety check: PassMIT2 mo ago
199

Zero token cost. An agent skill from LeoYeAI/openclaw-master-skills.

LeoYeAI/openclaw-master-skills2.2k—~889Automated safety check: PassMIT2 mo ago
200

Replay an LLM inference request trace (Mooncake / vLLM / SGLang hashids format) against a block-level KV prefix cache and compute hit statistics.

benchflow-ai/skillsbench1.8k—~2.3kAutomated safety check: PassApache-2.02 mo ago
201

Track real-time API cost accrual during LLM execution. An agent skill from curiositech/some_claude_skills.

curiositech/some_claude_skills2431 repo~2.3kAutomated safety check: PassMIT1 mo ago
202

Audit LLM token cost estimates against actual API usage. An agent skill from curiositech/some_claude_skills.

curiositech/some_claude_skills2431 repo~1.4kAutomated safety check: NotesMIT1 mo ago
203

Iterate on RAG systems with structured evals instead of eyeballing.

glebis/claude-skills389—~1.5kAutomated safety check: PassMIT11 days ago
204

Anthropic API prompt caching: TTL, breakpoints, stacking, invalidation, hit rate.

softspark/ai-toolkit179—~1.1kAutomated safety check: PassApache-2.0yesterday
205

Measure before optimizing — estimate token counts locally with stated heuristics, price them at your model's rates, and quantify before/after savings, because token optimization without measurement…

mohitagw15856/pm-claude-skills1.4k—~1.3kAutomated safety check: PassMITtoday
206

Analyzes markdown files for token efficiency. An agent skill from microsoft/GitHub-Copilot-for-Azure.

microsoft/GitHub-Copilot-for-Azure255—~364Automated safety check: PassMITtoday
207

Optimize Anthropic API costs — model selection, prompt caching, batches, Use when working with cost-tuning patterns.

jeremylongshore/tons-of-skills-marketplace2.8k—~1.3kAutomated safety check: PassMITtoday
208

Optimize Anthropic API latency — streaming, prompt caching, model selection, Use when working with performance-tuning patterns.

jeremylongshore/tons-of-skills-marketplace2.8k—~1.3kAutomated safety check: PassMITtoday
209

Analyze and control Flexport request volume without inventing undocumented global limits or headers.

jeremylongshore/tons-of-skills-marketplace2.8k—~1.1kAutomated safety check: PassMITtoday
210

A skill your agent uses when you need to keep PII out of Groq API calls, filter model responses, audit-log conversations, or track token cost and usage for a Groq integration.

jeremylongshore/tons-of-skills-marketplace2.8k—~1.3kAutomated safety check: NotesMITtoday
211

Set up observability for Groq integrations: latency histograms, token throughput, rate limit gauges, cost tracking, and Prometheus alerts.

jeremylongshore/tons-of-skills-marketplace2.8k—~1.6kAutomated safety check: PassMITtoday
212

Optimize context window usage for OpenRouter models to reduce cost and improve quality.

jeremylongshore/tons-of-skills-marketplace2.8k—~2.4kAutomated safety check: PassMITtoday
213

Implement comprehensive observability for LLM applications including tracing (Langfuse/Helicone), cost tracking, token optimization, RAG evaluation metrics (RAGAS), hallucination detection, and…

omer-metin/skills-for-antigravity162—~578Automated safety check: PassApache-2.08 mo ago
214

Context management engine for AI coding agents. An agent skill from borghei/Claude-Skills.

borghei/Claude-Skills881—~2.2kAutomated safety check: PassMITtoday
215

Every PostHog resource in one CLI — with offline search, agent-native output, and cross-resource analytics no...

mvanhorn/printing-press-library2.1k—~3.2kAutomated safety check: NotesApache-2.0today
216

Track Clawdbot AI model usage and estimate costs. An agent skill from sundial-org/awesome-openclaw-skills.

sundial-org/awesome-openclaw-skills663—~622Automated safety check: PassNo licence7 mo ago
217

Preserve conversation continuity across token compaction cycles by extracting and archiving all prompts with date-wise entries.

sundial-org/awesome-openclaw-skills663—~906Automated safety check: PassNo licence7 mo ago
218

Build an optimal context pack for the user's task — ranked file list with offset/limit suggestions, based on git state, mentioned paths, and historical patterns

egorfedorov/claude-context-optimizer114—~448Automated safety check: PassMIT8 days ago
219
219.Skill ReviewerOfficial

Review skill PRs with structured severity-rated feedback covering token budgets, routing conflicts, and repo conventions.

microsoft/GitHub-Copilot-for-Azure255—~482Automated safety check: PassMITtoday
220

A skill your agent uses when metering and capping AI or cloud app spend — tokens read from the response usage object, priced off a dated rate table, ledgered per user/tenant/feature, with alerts and…

ericrisco/rsc-harness167—~3.2kAutomated safety check: PassMITtoday
221

A skill your agent uses when calling open-weight LLMs on Together AI or Fireworks AI's OpenAI-compatible endpoints — baseurl plus namespaced model id, the cheapest model that clears the bar…

ericrisco/rsc-harness167—~3.3kAutomated safety check: PassMITtoday
222

Real-time cost tracking, budget enforcement, and ROI measurement for AI agent operations.

majiayu000/claude-skill-registry6661 repo~5kAutomated safety check: NotesMITtoday
223

Comprehensive cost tracking and optimization for production Claude deployments.

majiayu000/claude-skill-registry6661 repo~3.1kAutomated safety check: PassMITtoday
224

Compress selected context to a target token budget while preserving decisions, evidence, constraints, and unresolved questions.

majiayu000/claude-skill-registry6661 repo~2.2kAutomated safety check: PassMITtoday
225

Compact one or more Spec-Driven Development artifacts in place to reduce token cost on subsequent /speckit.

haru/redmine_ai_helper106—~1.8kAutomated safety check: PassMITtoday
226

Toggle a project-local concise-output directive that suppresses agent prose padding during SDD steps.

haru/redmine_ai_helper106—~1.9kAutomated safety check: PassMITtoday
227

Build a focused reading manifest for the next workflow step.

haru/redmine_ai_helper106—~1.3kAutomated safety check: PassMITtoday
228

Show token usage for every SDD artifact in the active feature, the projected context window for each upcoming phase, and the savings already realized by token-budget (full backups vs current files).

haru/redmine_ai_helper106—~1kAutomated safety check: PassMITtoday
229

LLM API 使用成本优化模式——基于任务复杂度的模型路由、预算跟踪、重试逻辑和提示词缓存. An agent skill from xu-xiang/everything-claude-code-zh.

xu-xiang/everything-claude-code-zh2k—~1.1kAutomated safety check: PassMIT7 mo ago
230

Clay workflows — reduce a workflow's credit and LLM cost via the CLI (clay workflows commands).

clay-run/agent-plugins130—~1.3kAutomated safety check: PassNo licencetoday
231

Turn JSON or PostgreSQL jsonb payloads into compact readable context for LLMs.

aiskillstore/marketplace430—~1kAutomated safety check: PassNo licencetoday
232

Audit and optimize OpenClaw token usage, cron job efficiency, and agent performance.

LeoYeAI/openclaw-master-skills2.2k—~4.5kAutomated safety check: NotesMIT2 mo ago
233

Optimize Anthropic Claude API costs with model routing, prompt caching, batching, and spend monitoring.

jeremylongshore/tons-of-skills-marketplace2.8k—~2.1kAutomated safety check: PassMITtoday
234

Optimize Claude API performance with prompt caching, model selection, streaming, and latency reduction techniques.

jeremylongshore/tons-of-skills-marketplace2.8k—~1.9kAutomated safety check: PassMITtoday
235

Call POST /chat/completions on Venice. An agent skill from veniceai/skills.

veniceai/skills143—~5.9kAutomated safety check: PassMIT2 days ago
236

Compresses artifacts for judge evaluation. An agent skill from closedloop-ai/claude-plugins.

closedloop-ai/claude-plugins122—~2.1kAutomated safety check: NotesApache-2.0today
237

A skill your agent uses when the user wants to estimate or predict the cost, token usage, or time of a task BEFORE it runs — "how much will this feature cost to build", "estimate the tokens for this…

Habitat-Thinking/ai-literacy-superpowers114—~5.4kAutomated safety check: PassUnknown17 days ago
238

Every Groq endpoint in your terminal, plus a local ledger that tracks token cost and rate-limit budget.

mvanhorn/printing-press-library2.1k—~7.9kAutomated safety check: NotesApache-2.0today
239

Review what an LLM feature or agent actually puts in its context window — and find what's bloating, missing, or fighting itself.

mohitagw15856/pm-claude-skills1.4k—~1.4kAutomated safety check: PassMITtoday
240

Model the cost and latency of an LLM feature before it ships and surprises the bill.

mohitagw15856/pm-claude-skills1.4k—~985Automated safety check: PassMITtoday