Search

LLM cost and token optimization

254 skills found, page 3.
Search results
#SkillRepositoryStarsUsed inTokensAuto-checkLicenceUpdated
97

Controls the RTK filter set and context-handling settings in OmniRoute, with endpoints to try compression on sample text and read back retained output.

diegosouzapw/OmniRoute75k—~618Automated safety check: PassMITtoday
98

Queries OmniRoute call logs, usage history and analytics, filters them by provider, model, status or cost, and reads or sets usage budgets.

diegosouzapw/OmniRoute75k—~2kAutomated safety check: PassMITtoday
99

Verifies how Cline marks Anthropic prompt-cache breakpoints on the wire, then adds a tools breakpoint and tunes the cache TTL in its AI SDK code.

OnlyTerp/prompt-cache-skills114—~2kAutomated safety check: PassUnknown1 mo ago
100

Configure token budget limits, auto-compact settings, and view current budget status (model-aware — Claude 5 lineup, Opus 5.5 default fallback, full 1M context at standard price)

egorfedorov/claude-context-optimizer114—~808Automated safety check: PassMIT11 days ago
101

Builds an explorable HTML report of Claude Code usage over a chosen period: tokens, cache behavior, subagents, skills and costly prompts.

anthropics/claude-plugins-official38k—~784Automated safety check: PassApache-2.0today
102

Always-on context compression for OpenClaw. An agent skill from Supercompress/Supercompress.

Supercompress/Supercompress107—~366Automated safety check: PassMITyesterday
103

Estimate Kubernetes infrastructure costs by querying cluster node, pod, PVC/PV, and LoadBalancer data, applying cloud pricing models, and producing cost attribution reports with storage and…

initializ/forge222—~2.7kAutomated safety check: PassApache-2.02 days ago
104

Check token budget and run-log spend before and after a loop run. Enforces early exit when over budget or when there is no actionable work.

cobusgreyling/loop-engineering11k1 repo~376Automated safety check: PassMITtoday
105

Patch guide for adding a stable per-task prompt_cache_key to Cline's OpenAI native provider so cached token counts stop reading as zero.

OnlyTerp/prompt-cache-skills114—~868Automated safety check: PassUnknown1 mo ago
106

This skill should be used for improving context efficiency: context budgeting, observation masking, prefix or KV-cache strategy, partitioning, token-cost reduction, retrieval scoping, and extending…

guanyang/open-agent-hub9772 repos~4kAutomated safety check: PassMITtoday
107

This skill should be used for persistent semantic memory in agent systems: cross-session knowledge retention, entity tracking, temporal validity, graph or vector retrieval, memory consolidation, and…

guanyang/open-agent-hub9772 repos~4.1kAutomated safety check: PassMITtoday
108

Designs, tests and refines LLM prompts: zero-shot, few-shot and chain-of-thought patterns, system prompts, structured output schemas and evaluation test suites.

Jeffallan/claude-skills12k—~1.5kAutomated safety check: PassMIT7 days ago
109

An advanced skill for L3 autonomous loops. An agent skill from cobusgreyling/loop-engineering.

cobusgreyling/loop-engineering11k1 repo~864Automated safety check: PassMITtoday
110

Add a new provider API capability (prompt caching, strict/structured tool calling, thinking/reasoning effort, service tier, safety settings, logprobs, etc.) to Pydantic AI.

pydantic/pydantic-ai21k—~3.2kAutomated safety check: PassMITtoday
111

Prepares many-file, single-turn transforms such as translating or rewriting as a plan, then submits it to the asynchronous, half-price DashScope Batch API through the qwen batch CLI.

QwenLM/qwen-code28k—~2.2kAutomated safety check: PassApache-2.0today
112

Cache the parts of the prompt that don't change so a long-running loop stops paying full price on every turn.

Archive228/loopkit755—~735Automated safety check: PassMIT2 mo ago
113

A skill your agent uses when the user wants to capture AI tool costs, review spending trends, set cost budgets, or integrate cost data into health snapshots — guides quarterly cost capture, records…

Habitat-Thinking/ai-literacy-superpowers114—~1.7kAutomated safety check: PassUnknown20 days ago
114

Specifies how quota-axi reads LLM subscription quota windows, derives pace and runway, computes a selection signal and renders output in TOON, JSON and TUI tiers.

kunchenguid/quota-axi147—~1.2kAutomated safety check: PassMITyesterday
115

Caching strategies for LLM prompts including Anthropic prompt caching, response caching, and CAG (Cache Augmented Generation) Use when: prompt caching, cache prompt, response cache, cag, cache…

davila7/claude-code-templates33k5 repos~452Automated safety check: PassMITtoday
116

Guide for debugging OhMyCode issues. An agent skill from AlphaLab-USTC/OhMyCode.

AlphaLab-USTC/OhMyCode131—~1.3kAutomated safety check: PassMIT6 mo ago
117

Checks every subagent prompt before spawning, swapping pasted files and context for paths and short summaries and trimming the brief, to avoid multiplied token cost.

LichAmnesia/lich-skills234—~1.6kAutomated safety check: PassMIT4 mo ago
118

追蹤並計算當前 session 的 Token 使用量和花費(USD)。產出前後對比報告,幫助使用者了解 workspace 優化的實際效益。

zeuikli/claude-code-workspace156—~830Automated safety check: PassNo licence5 days ago
119

Track and report Claude Code token usage, spending, and budgets from the local ECC cost-tracker metrics log.

affaan-m/ECC277k1 repo~1.3kAutomated safety check: PassMITtoday
120

Continue's prompt caching is opt-in via config and off by default.

OnlyTerp/prompt-cache-skills114—~977Automated safety check: PassUnknown1 mo ago
121

Track session costs, set budget alerts, and optimize token spend.

rohitg00/pro-workflow2.9k—~814Automated safety check: PassNo licence12 days ago
122

A skill your agent uses when the user asks to optimize prompts, design prompt templates, evaluate LLM outputs with an eval set, measure RAG retrieval quality, validate agent/tool configurations…

alirezarezvani/claude-skills28k1 repo~2.5kAutomated safety check: PassMIT1 mo ago
123

Generate a concise daily infrastructure briefing. An agent skill from bolivian-peru/os-moda.

bolivian-peru/os-moda119—~897Automated safety check: PassApache-2.03 mo ago
124

A skill your agent uses to classify the documentation impact of a pull request diff, returning one of three verdicts -- no-change, in-place edit, or structural change -- with bounded LLM cost.

microsoft/apm4k—~1.8kAutomated safety check: PassMITyesterday
125

This skill should be used when writing, enhancing, or evaluating the launch prompt for a long-running autonomous agent or a parallel multi-agent orchestration attacking a hard problem: pseudo-formal…

guanyang/open-agent-hub9771 repo~6.4kAutomated safety check: PassMITtoday
126

This skill should be used when a model gets read-write control over its own live context window instead of a harness-scheduled compaction policy: the context exposed as an editable file the model…

guanyang/open-agent-hub9771 repo~6kAutomated safety check: PassMITtoday
127
127.Cost

AI Build Cost Tracker — track how much AI is costing you per feature.

Houseofmvps/ultraship123—~1.1kAutomated safety check: NotesMIT3 mo ago
128

A skill your agent uses when context compression is imminent, when resuming a session, or when preserving critical decisions across long tasks.

vibeeval/vibecosystem532—~1.7kAutomated safety check: PassMIT2 mo ago
129

Context window management and token optimization guide. An agent skill from jh941213/my-cc-harness.

jh941213/my-cc-harness125—~621Automated safety check: PassNo licence2 mo ago
130

Analyze markdown files for token efficiency and reduce context-window bloat.

Azure/azure-sdk-tools134—~602Automated safety check: PassMITyesterday
131
131.App ObservabilityOfficial

Get RED metrics + service maps + frontend RUM + AI/LLM monitoring out of Grafana Cloud — Application Observability (tracesspanmetrics from OTel traces, p50/p95/p99 latency, exemplar-to-trace…

grafana/skills282—~1.8kAutomated safety check: PassApache-2.02 days ago
132

Analyze and reduce token consumption in agentic workflows — guardrail-specific entry points, measurement, and optimization techniques.

github/gh-aw5.4k—~1.2kAutomated safety check: PassMITtoday
133

Monitor LLMs and agentic apps: performance, token/cost, response quality, and workflow orchestration.

aspectrr/deer405—~858Automated safety check: PassMIT5 mo ago
134

Create and manage webhook subscriptions for event-driven agent activation, or for direct push notifications (zero LLM cost).

Tommy-yw/RunbookHermes5461 repo~1.7kAutomated safety check: NotesMIT4 mo ago
135

Orchestrate autonomous AI development pipelines through your Kanban board (Asana, GitHub Projects, Linear).

sickn33/agentic-awesome-skills47k2 repos~2kAutomated safety check: PassMIT2 days ago
136

Implement multi-layer LLM caching with exact match, semantic similarity, and provider-side prompt caching.

sickn33/agentic-awesome-skills47k2 repos~2.8kAutomated safety check: PassMIT2 days ago
137

Caching strategies for LLM prompts including Anthropic prompt caching, response caching, and CAG (Cache Augmented Generation)

sickn33/agentic-awesome-skills47k2 repos~3.4kAutomated safety check: PassMIT2 days ago
138

OpenCode places cachePoint on Bedrock messages containing DocumentBlocks, which produces a "nothing available to cache" error.

OnlyTerp/prompt-cache-skills114—~966Automated safety check: PassUnknown1 mo ago
139

Optimize LLM prompt caching hit rate to reduce API costs and improve latency.

affaan-m/ECC276k—~2.5kAutomated safety check: PassMITyesterday
140

Ranked symbol map of a codebase within a token budget — a compact "what matters in this repo" before reading files.

AnastasiyaW/codex-claude-code-config154—~1.5kAutomated safety check: PassMITyesterday
141

Master the four operations of context engineering — Write, Select, Compress, Isolate.

rohitg00/pro-workflow2.9k—~1.6kAutomated safety check: PassNo licence12 days ago
142

Send images, audio, video, or documents into an AG2 beta Agent alongside text.

ag2ai/build-with-ag2252—~1.7kAutomated safety check: PassApache-2.01 mo ago
143

Investigate LLM spend in PostHog — total cost over time, cost by model, provider, user, trace, or custom dimension, token and cache-hit economics, and cost regressions.

PostHog/posthog40k—~2.9kAutomated safety check: PassUnknownyesterday
144

Orchestrate free opencode workers from Claude Code to cut token costs.

alirezarezvani/claude-skills28k—~2.4kAutomated safety check: PassApache-2.01 mo ago