Search

LLM cost and token optimization

254 skills found, page 5.
Search results
#SkillRepositoryStarsUsed inTokensAuto-checkLicenceUpdated
193

Analyze & debug GenAI/LLM apps: token cost & caching by prompt, model & provider; latency/errors; agent & tool loops/failures; conversations; guardrails; evaluations; OpenTelemetry/dt-evals setup.

Dynatrace/dynatrace-for-ai163—~4.5kAutomated safety check: PassApache-2.010 days ago
194

A skill your agent uses when building or debugging apps that call the Claude API — implementing tool use, streaming, vision, prompt caching, batch processing, extended thinking, or an agentic loop…

kid-sid/claude-spellbook190—~2.7kAutomated safety check: PassMIT2 mo ago
195

Reduce OpenClaw token usage and API costs through smart model routing, heartbeat optimization, budget tracking, and native 2026.2.15 features (session pruning, bootstrap size limits, cache TTL…

LeoYeAI/openclaw-master-skills2.2k—~5.2kAutomated safety check: PassMIT2 mo ago
196

Zero token cost. An agent skill from LeoYeAI/openclaw-master-skills.

LeoYeAI/openclaw-master-skills2.2k—~889Automated safety check: PassMIT2 mo ago
197

A skill your agent uses for debugging DSPy programs, inspecthistory, tracing LLM calls, custom callbacks, observability, monitoring, and cost tracking.

OmidZamani/dspy-skills123—~2.1kAutomated safety check: WarnMIT3 mo ago
198

Replay an LLM inference request trace (Mooncake / vLLM / SGLang hashids format) against a block-level KV prefix cache and compute hit statistics.

benchflow-ai/skillsbench1.8k—~2.3kAutomated safety check: PassApache-2.02 mo ago
199

Iterate on RAG systems with structured evals instead of eyeballing.

glebis/claude-skills391—~1.5kAutomated safety check: PassMIT3 days ago
200

Anthropic API prompt caching: TTL, breakpoints, stacking, invalidation, hit rate.

softspark/ai-toolkit179—~1.1kAutomated safety check: PassApache-2.03 days ago
201

Measure before optimizing — estimate token counts locally with stated heuristics, price them at your model's rates, and quantify before/after savings, because token optimization without measurement…

mohitagw15856/pm-claude-skills1.4k—~1.3kAutomated safety check: PassMIT2 days ago
202

Analyzes markdown files for token efficiency. An agent skill from microsoft/GitHub-Copilot-for-Azure.

microsoft/GitHub-Copilot-for-Azure255—~364Automated safety check: PassMITyesterday
203

Monitor and optimize LLM costs using Langfuse analytics and dashboards.

jeremylongshore/tons-of-skills-marketplace2.8k—~2.4kAutomated safety check: PassMITyesterday
204

Optimize Anthropic API costs — model selection, prompt caching, batches, Use when working with cost-tuning patterns.

jeremylongshore/tons-of-skills-marketplace2.8k—~1.3kAutomated safety check: PassMITyesterday
205

Optimize Anthropic API latency — streaming, prompt caching, model selection, Use when working with performance-tuning patterns.

jeremylongshore/tons-of-skills-marketplace2.8k—~1.3kAutomated safety check: PassMITyesterday
206

Analyze and control Flexport request volume without inventing undocumented global limits or headers.

jeremylongshore/tons-of-skills-marketplace2.8k—~1.1kAutomated safety check: PassMITyesterday
207

A skill your agent uses when you need to keep PII out of Groq API calls, filter model responses, audit-log conversations, or track token cost and usage for a Groq integration.

jeremylongshore/tons-of-skills-marketplace2.8k—~1.3kAutomated safety check: NotesMITyesterday
208

Set up observability for Groq integrations: latency histograms, token throughput, rate limit gauges, cost tracking, and Prometheus alerts.

jeremylongshore/tons-of-skills-marketplace2.8k—~1.6kAutomated safety check: PassMITyesterday
209

Optimize context window usage for OpenRouter models to reduce cost and improve quality.

jeremylongshore/tons-of-skills-marketplace2.8k—~2.4kAutomated safety check: PassMITyesterday
210

Implement comprehensive observability for LLM applications including tracing (Langfuse/Helicone), cost tracking, token optimization, RAG evaluation metrics (RAGAS), hallucination detection, and…

omer-metin/skills-for-antigravity163—~578Automated safety check: PassApache-2.08 mo ago
211

Context management engine for AI coding agents. An agent skill from borghei/Claude-Skills.

borghei/Claude-Skills891—~2.2kAutomated safety check: PassMIT4 days ago
212

Every PostHog resource in one CLI — with offline search, agent-native output, and cross-resource analytics no...

mvanhorn/printing-press-library2.1k—~3.2kAutomated safety check: NotesApache-2.0yesterday
213

Build an optimal context pack for the user's task — ranked file list with offset/limit suggestions, based on git state, mentioned paths, and historical patterns

egorfedorov/claude-context-optimizer114—~448Automated safety check: PassMIT11 days ago
214

A skill your agent uses when metering and capping AI or cloud app spend — tokens read from the response usage object, priced off a dated rate table, ledgered per user/tenant/feature, with alerts and…

ericrisco/rsc-harness180—~3.2kAutomated safety check: PassMITyesterday
215

A skill your agent uses when calling open-weight LLMs on Together AI or Fireworks AI's OpenAI-compatible endpoints — baseurl plus namespaced model id, the cheapest model that clears the bar…

ericrisco/rsc-harness180—~3.3kAutomated safety check: PassMITyesterday
216

Track Clawdbot AI model usage and estimate costs. An agent skill from sundial-org/awesome-openclaw-skills.

sundial-org/awesome-openclaw-skills663—~622Automated safety check: PassNo licence7 mo ago
217

Preserve conversation continuity across token compaction cycles by extracting and archiving all prompts with date-wise entries.

sundial-org/awesome-openclaw-skills663—~906Automated safety check: PassNo licence7 mo ago
218
218.Skill ReviewerOfficial

Review skill PRs with structured severity-rated feedback covering token budgets, routing conflicts, and repo conventions.

microsoft/GitHub-Copilot-for-Azure255—~482Automated safety check: PassMITyesterday
219

Compact one or more Spec-Driven Development artifacts in place to reduce token cost on subsequent /speckit.

haru/redmine_ai_helper106—~1.8kAutomated safety check: PassMITyesterday
220

Toggle a project-local concise-output directive that suppresses agent prose padding during SDD steps.

haru/redmine_ai_helper106—~1.9kAutomated safety check: PassMITyesterday
221

Build a focused reading manifest for the next workflow step.

haru/redmine_ai_helper106—~1.3kAutomated safety check: PassMITyesterday
222

Show token usage for every SDD artifact in the active feature, the projected context window for each upcoming phase, and the savings already realized by token-budget (full backups vs current files).

haru/redmine_ai_helper106—~1kAutomated safety check: PassMITyesterday
223

LLM API 使用成本优化模式——基于任务复杂度的模型路由、预算跟踪、重试逻辑和提示词缓存. An agent skill from xu-xiang/everything-claude-code-zh.

xu-xiang/everything-claude-code-zh2k—~1.1kAutomated safety check: PassMIT7 mo ago
224

Clay workflows — reduce a workflow's credit and LLM cost via the CLI (clay workflows commands).

clay-run/agent-plugins132—~1.3kAutomated safety check: PassNo licence3 days ago
225

Turn JSON or PostgreSQL jsonb payloads into compact readable context for LLMs.

aiskillstore/marketplace433—~1kAutomated safety check: PassNo licenceyesterday
226

Audit and optimize OpenClaw token usage, cron job efficiency, and agent performance.

LeoYeAI/openclaw-master-skills2.2k—~4.5kAutomated safety check: NotesMIT2 mo ago
227

Optimize Anthropic Claude API costs with model routing, prompt caching, batching, and spend monitoring.

jeremylongshore/tons-of-skills-marketplace2.8k—~2.1kAutomated safety check: PassMITyesterday
228

Optimize Claude API performance with prompt caching, model selection, streaming, and latency reduction techniques.

jeremylongshore/tons-of-skills-marketplace2.8k—~1.9kAutomated safety check: PassMITyesterday
229

Reduce LLM API and infrastructure costs through model selection, prompt caching, batching, caching, quantization, and self-hosting strategies.

BagelHole/DevOps-Security-Agent-Skills1.2k—~2.2kAutomated safety check: PassMIT4 mo ago
230

Call POST /chat/completions on Venice. An agent skill from veniceai/skills.

veniceai/skills144—~5.9kAutomated safety check: PassMIT5 days ago
231

Compresses artifacts for judge evaluation. An agent skill from closedloop-ai/claude-plugins.

closedloop-ai/claude-plugins122—~2.1kAutomated safety check: NotesApache-2.0yesterday
232

A skill your agent uses when the user wants to estimate or predict the cost, token usage, or time of a task BEFORE it runs — "how much will this feature cost to build", "estimate the tokens for this…

Habitat-Thinking/ai-literacy-superpowers114—~5.4kAutomated safety check: PassUnknown20 days ago
233

Every Groq endpoint in your terminal, plus a local ledger that tracks token cost and rate-limit budget.

mvanhorn/printing-press-library2.1k—~7.9kAutomated safety check: NotesApache-2.0yesterday
234

Review what an LLM feature or agent actually puts in its context window — and find what's bloating, missing, or fighting itself.

mohitagw15856/pm-claude-skills1.4k—~1.4kAutomated safety check: PassMIT2 days ago
235

Model the cost and latency of an LLM feature before it ships and surprises the bill.

mohitagw15856/pm-claude-skills1.4k—~985Automated safety check: PassMIT2 days ago
236

Choose the right LLM for a task by trading off quality, cost, latency, and constraints.

mohitagw15856/pm-claude-skills1.4k—~1.1kAutomated safety check: PassMIT2 days ago
237

Cut LLM output tokens 40–70% by stripping grammatical scaffolding while preserving every fact — telegraphic output modes, when they pay (pipelines, long sessions) and when they don't (single shots…

mohitagw15856/pm-claude-skills1.4k—~1.5kAutomated safety check: PassMIT2 days ago
238

Ultimate cost optimization toolkit for OpenClaw/Claude Code.

LeoYeAI/openclaw-master-skills2.2k—~3.3kAutomated safety check: NotesMIT2 mo ago
239

Java/Maven single-module deep documentation generator. An agent skill from LeoYeAI/openclaw-master-skills.

LeoYeAI/openclaw-master-skills2.2k—~3.4kAutomated safety check: PassMIT2 mo ago
240

A skill your agent uses for direct LiteLLM Python SDK work: chat/text completions, async calls, streaming, embeddings, structured outputs, tools, token/cost checks, caching, callbacks, import/smoke…

VectorSpaceLab/AREX-Skill331—~1.1kAutomated safety check: PassUnknown1 mo ago