Topic · AI & LLM Engineering
Best LLM cost and token optimization skills, page 2
LLM cost and token optimization skills, ranked
Ranked by score. Sort bymost stars,trending,newest,recently updated
| # | Skill | Repository | Stars | Used in | Tokens | Auto-check | Licence | Updated |
|---|---|---|---|---|---|---|---|---|
| 49 | Check token budget and run-log spend before and after a loop run. Enforces early exit when over budget or when there is no actionable work. | cobusgreyling/ | 11k | 1 repo | ~376 | Automated safety check: Pass | MIT | today |
| 50 | 50.AI A skill your agent uses when calling the app's AI gateway from agent tools — chat completions, embeddings, listing models, configuring defaults or BYOK, reading token/cost usage | butterbase-ai/ | 534 | — | ~1.1k | Automated safety check: Pass | MIT | 2 days ago |
| 51 | 51.Honey Write less code and say less about it. An agent skill from Green-PT/honey-for-devs. | Green-PT/ | 310 | — | ~3.5k | Automated safety check: Pass | MIT | 1 mo ago |
| 52 | 分析 kiro2cc-proxy 访问日志,计算 Prompt Caching 节省的 credits。只要用户粘贴了含有"输入token 输出token 费用$ credits✓"格式的日志行,并询问节省了多少credits、缓存效率、cost分析等,立即使用此 skill。触发关键词:节省了多少credits、cache节省、分析日志、caching… | TsinHzl/ | 163 | — | ~1.1k | Automated safety check: Pass | MIT | today |
| 53 | Plan, smoke-test, execute, checkpoint, publish, audit, and reproduce full ShellBench native benchmark campaigns across OpenClaw, Hermes, Codex, and Claude Code, including model and reasoning… | openclaw/ | 141 | — | ~1.1k | Automated safety check: Pass | MIT | yesterday |
| 54 | Run the weak-agent adversarial test harness against docx-cli. | kklimuk/ | 215 | — | ~6.1k | Automated safety check: Notes | MIT | 12 days ago |
| 55 | 55.Vibe Delegate a coding task to Mistral Vibe and supervise the result via git diff. | pcx-wave/ | 119 | — | ~7.2k | Automated safety check: Notes | MIT | 6 days ago |
| 56 | Mines local Claude Code session transcripts with a deterministic Python pipeline to show what the agent is actually used for, how often it fails and what it costs. | amd/ | 1.6k | — | ~2.3k | Automated safety check: Pass | MIT | yesterday |
| 57 | Surgically select, compress, and recover codebase context using Entroly's MCP tools. | juyterman1000/ | 472 | — | ~501 | Automated safety check: Pass | Apache-2.0 | today |
| 58 | 58.Wiki Core Core operating rules for the student knowledge wiki. An agent skill from IssacW228/student-llm-wiki. | IssacW228/ | 176 | — | ~708 | Automated safety check: Pass | MIT | 3 mo ago |
| 59 | Estimates the per-turn token cost of a project's .claude folder and CLAUDE.md, split into always-loaded, path-scoped and invoked-only files, and flags what runs over budget. | poshan0126/ | 871 | — | ~1.4k | Automated safety check: Pass | MIT | 1 mo ago |
| 60 | A skill your agent uses for genuinely high-volume Levyra work such as builds, tests, lint, logs, broad searches, dependency output, Git/GitHub or CodeRabbit inspection, CI diagnostics, agent setup… | LUC4N3X/ | 543 | — | ~1.3k | Automated safety check: Notes | GPL-3.0 | today |
| 61 | Audits a project's Claude Code setup for common token-cost leaks and returns a prioritized fix list, using only static inspection of files and settings. | mergisi/ | 4k | — | ~1.1k | Automated safety check: Pass | MIT | 11 days ago |
| 62 | 62.Review Delta Reviews only the code changed since the last commit, using a code-graph MCP server to find risk-scored changes, test gaps and the blast radius, then writes a short report. | tirth8205/ | 32k | — | ~325 | Automated safety check: Pass | MIT | yesterday |
| 63 | Queries Langfuse traces, prompts, datasets and sessions, and analyzes local LLM gateway logs for requests, context growth, token use and cache hits. | KonghaYao/ | 223 | — | ~4.3k | Automated safety check: Notes | Apache-2.0 | today |
| 64 | 64.Token Doctor Personal diagnosis of where your Claude Code + Cowork spend goes. | techwolf-ai/ | 132 | — | ~4.1k | Automated safety check: Pass | MIT | 8 days ago |
| 65 | Benchmarks AMD's GAIA agent against Claude Code and across models on quality, honesty, steps, tokens, time and real cost, using gaia eval tasks. | amd/ | 1.6k | — | ~1.8k | Automated safety check: Pass | MIT | yesterday |
| 66 | Use Entroly for content-blind agent-history audits, explicit baseline/optimized trials, recoverable command or browser evidence, response contracts, and token-efficiency claim verification. | juyterman1000/ | 472 | — | ~528 | Automated safety check: Pass | Apache-2.0 | today |
| 67 | Rewrites a memory file such as CLAUDE.md or a todo list in terse caveman-style text to cut input tokens, saving a readable backup outside the project tree. | JuliusBrussee/ | 940 | — | ~1.2k | Automated safety check: Pass | MIT | 1 mo ago |
| 68 | Split a multi-call LM workflow by cognitive load, not by accuracy: let one strong model make the few reasoning decisions and a cheap model do the many mechanical executions (Aider architect+editor… | agentsope/ | 457 | — | ~3k | Automated safety check: Pass | MIT | 1 mo ago |
| 69 | Routes a turn or delegated task to the cheapest model and effort lane that will still do it right, using the Jev decision model to classify difficulty and escalate only when needed. | kerpopule/ | 1k | — | ~2.7k | Automated safety check: Pass | MIT | yesterday |
| 70 | Always-on context compression for Grok Build. An agent skill from Supercompress/Supercompress. | Supercompress/ | 106 | — | ~519 | Automated safety check: Pass | MIT | 2 days ago |
| 71 | 71.Authoring Add or edit an MCP server entry in mcptoon's config correctly — stdio/streamable-http/sse shapes, command rules, placeholders, and validation. | activeing123/ | 214 | — | ~559 | Automated safety check: Pass | Apache-2.0 | today |
| 72 | 72.Paper Reader Deep-read an arXiv paper and generate structured Deep Note reading notes. | AlphaLab-USTC/ | 134 | — | ~1.6k | Automated safety check: Pass | MIT | 6 mo ago |
| 73 | Audit your OpenClaw setup for token waste, context bloat, and cost optimization opportunities | alexgreensh/ | 2.5k | — | ~481 | Automated safety check: Pass | Unknown | 2 days ago |
| 74 | 74.Token Saver Minimize token consumption & maximize prompt cache hit rate. | lokikill123/ | 119 | — | ~298 | Automated safety check: Pass | MIT | 3 mo ago |
| 75 | Describes ClawRouter, a local proxy that forwards each LLM request to the blockrun.ai gateway, which routes to a cheaper capable model, paid by USDC wallet or API key credit. | BlockRunAI/ | 6.6k | — | ~6.8k | Automated safety check: Pass | MIT | 2 days ago |
| 76 | Speeds up or throttles cognee ingestion: estimate cost with a dry run, tune batching and chunk size, set LLM rate limits and run work in the background. | topoteretes/ | 32k | — | ~2.5k | Automated safety check: Pass | Apache-2.0 | today |
| 77 | Audit and remediate Entroly's MCP marketplace quality with evidence, adversarial validation, and no score gaming. | juyterman1000/ | 472 | — | ~1.2k | Automated safety check: Pass | Apache-2.0 | today |
| 78 | Starts a new Claude Code session from only the last quarter of the current conversation, dropping earlier context to save tokens while keeping recent work. | ykdojo/ | 10k | — | ~401 | Automated safety check: Pass | Unknown | 12 days ago |
| 79 | The reference agents' cache-stable request assembly, covering the static system and per-request context split, the fixed tool list, the rolling conversation breakpoint, which config fields are… | anthropics/ | 3.2k | — | ~2k | Automated safety check: Pass | Apache-2.0 | 5 days ago |
| 80 | 80.Clawmetry Real-time observability for OpenClaw agents — local dashboard + optional encrypted cloud sync. | vivekchand/ | 424 | — | ~992 | Automated safety check: Pass | MIT | today |
| 81 | Shows today's Claude Code spending through the claude-view MCP server, with total cost, running sessions and a per-session breakdown, and other date ranges on request. | tombelieber/ | 110 | — | ~3k | Automated safety check: Pass | MIT | 7 days ago |
| 82 | 82.Triage Diagnose and fix broken MCP servers in mcptoon — doctor, health, per-server probes, and the common failure playbook. | activeing123/ | 214 | — | ~488 | Automated safety check: Pass | Apache-2.0 | today |
| 83 | Shows a markdown dashboard of basemind activity in the session: tool call counts, top operations and estimated tokens saved against a grep and Read baseline. | Goldziher/ | 106 | — | ~558 | Automated safety check: Pass | MIT | yesterday |
| 84 | Patches Aider to request Anthropic's one-hour prompt cache lifetime, so long pauses between turns no longer expire the cache or need paid keepalive pings. | OnlyTerp/ | 114 | — | ~990 | Automated safety check: Pass | Unknown | 1 mo ago |
| 85 | Teaches how xc-plugin saves tokens with progressive disclosure, cached responses and consistent configuration, so large device lists and build logs arrive as summaries first. | conorluddy/ | 183 | — | ~4k | Automated safety check: Pass | MIT | 25 days ago |
| 86 | Trigger when the user asks which model to use, wants to compare model costs, says "what's cheapest for this task", "should I use Opus or Sonnet", "can a smaller model handle this", or… | mergisi/ | 4k | — | ~1k | Automated safety check: Pass | MIT | 11 days ago |
| 87 | 87.AI Agent Quick reference for the Breeze RMM AI Agent system architecture, MCP tools, streaming chat, cost tracking, guardrails, and MCP server. | LanternOps/ | 130 | — | ~3.7k | Automated safety check: Pass | AGPL-3.0 | today |
| 88 | Ranks a large catalog of installed skills against the current request through the Jev service, and can conclude that no skill applies. | kerpopule/ | 1k | — | ~2.2k | Automated safety check: Pass | MIT | yesterday |
| 89 | Configure and test prompt compression from the CLI. Manage RTK filters, Caveman rules, stacked compression modes, and preview compression output… | diegosouzapw/ | 74k | — | ~521 | Automated safety check: Pass | MIT | yesterday |
| 90 | Analyze and safely optimize OpenSpec main specifications. An agent skill from cdavid817/vanehub-ai. | cdavid817/ | 116 | — | ~347 | Automated safety check: Pass | Apache-2.0 | yesterday |
| 91 | Chooses which paid frontier model seat should take a task already judged hard, hands it off with proper context, and keeps a watch on the delegated run. | kerpopule/ | 1k | — | ~1.5k | Automated safety check: Warn | MIT | yesterday |
| 92 | Patch recipe that turns on Aider's --cache-prompts by default for models that support caching, with a verification procedure against the wire request. | OnlyTerp/ | 114 | — | ~638 | Automated safety check: Pass | Unknown | 1 mo ago |
| 93 | 93.Memory Persistent global memory across conversations. An agent skill from lokikill123/codex-token-skills. | lokikill123/ | 119 | — | ~304 | Automated safety check: Pass | MIT | 3 mo ago |
| 94 | Lists which models a Weave Router installation may route to and turns individual models or whole providers on or off, mirroring the router dashboard's settings. | weave-os/ | 5.6k | — | ~533 | Automated safety check: Pass | Apache-2.0 | today |
| 95 | This skill should be used when creating conversation summaries for CodeMemory, discussing summarization quality, optimizing DAG structure, or working with the compaction system. | harrylettering/ | 158 | — | ~788 | Automated safety check: Pass | No licence | 2 days ago |
| 96 | Manage context engineering configurations, RTK filter sets, and conversation sessions from the CLI. Apply context-relay settings and inspect active… | diegosouzapw/ | 74k | — | ~1.2k | Automated safety check: Pass | MIT | yesterday |
Explore related skills
Category
More topics in AI & LLM Engineering
- Building AI agents525
- Deep learning408
- Embeddings381
- LLM inference and serving364
- Retrieval-augmented generation360
- Prompt engineering350
- Fine-tuning313
- LLM evaluation303
- Speech recognition and synthesis272
- Structured output and tool calling271
- LLM API integration218
- LLM observability217
- Model routing and gateways217
- LLM guardrails208
- Computer vision206
- Model hubs and datasets180
- GPU and accelerator computing171
- Diffusion and image models167
- Natural language processing141
- Reinforcement learning67
- AI interpretability23