Agent skill

OmniRoute LLM Cache

by diegosouzapw in diegosouzapw/OmniRoute

Documents OmniRoute's cache endpoints for reading cache statistics and clearing entries, statistics or the reasoning cache, with notes on TTL and similarity settings.

MITAuto-check passedBackend & APIs

Install OmniRoute LLM Cache

skills CLI
$ npx skills add diegosouzapw/OmniRoute --skill omni-cache -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install diegosouzapw/OmniRoute omni-cache --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/diegosouzapw/OmniRoute.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/omni-cache .claude/skills/omni-cache && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
omni-cache
GitHub stars
74k
Used in
1 other repo
Token cost
~529 tokens
SKILL.md length
106 words
Files
1
Skills in repo
50
Repo updated
First seen
Licence
MIT

At a glance

Documents OmniRoute's cache endpoints for reading cache statistics and clearing entries, statistics or the reasoning cache, with notes on TTL and similarity settings.

  • Checking how well the LLM response cache is performing
  • SKILL.md covers Overview, Authentication, Endpoints and Payloads
  • Calls curl; needs OMNIROUTE_TOKEN and REQUIRE_API_KEY
  • Clearing stale cached responses after changing a prompt or model

What it does

The endpoints shown work on several cache layers. A GET on /api/cache returns cache statistics and a DELETE clears all caches, /api/cache/stats returns detailed statistics for every layer and can clear them, /api/cache/entries lists or deletes entries, and /api/cache/reasoning does the same for reasoning data. Each is a curl example against a local server.

The description adds TTL policies and semantic-similarity thresholds, but the endpoints visible in the excerpt only read and clear. Calls need a Bearer token or session cookie, obtained from POST /api/auth/login, or REQUIRE_API_KEY set to false for local development, and the full schemas are in the OpenAPI spec.

When your agent uses it

  • Checking how well the LLM response cache is performing
  • Clearing stale cached responses after changing a prompt or model
  • Wiping the reasoning cache without touching other layers
  • Reviewing which entries the cache currently holds

Example prompts

  • “Show me the detailed OmniRoute cache statistics for every layer.”
  • “Clear all cached entries in OmniRoute but keep the statistics.”
  • “Delete only the reasoning cache on my local OmniRoute server.”

Requirements

  • A running OmniRoute server and a Bearer token or session cookie

What it can do on your machine

Read from SKILL.md and the folder at commit 8ad6b1c. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Shell commands in SKILL.md call:

    • curl

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md. Its commands use curl, which can reach the network depending on how they are called.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names these keys or tokens, usually read from environment variables:

    • OMNIROUTE_TOKEN
    • REQUIRE_API_KEY

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

OmniRoute LLM Cache loads about 529 tokens when it runs. Until then it costs about 39 tokens; SKILL.md has 106 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~39
When it runs · the whole SKILL.md, loaded when a task matches
~529

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from diegosouzapw/OmniRoute at commit 8ad6b1c, republished under its MIT licence (© diegosouzapw). 106 words, ~529 tokens.

Download SKILL.mdSave it as .claude/skills/omni-cache/SKILL.md (or your agent's skills folder).
name
omni-cache
description
Manage the LLM response cache. View cache statistics, clear entries, configure TTL policies, and control semantic-similarity caching thresholds.
<!-- generated by src/lib/agentSkills/generator.ts; manual edits will be overwritten -->

Overview

Manage the LLM response cache. View cache statistics, clear entries, configure TTL policies, and control semantic-similarity caching thresholds.

Authentication

All requests require a valid Bearer token or session cookie. Obtain a token via POST /api/auth/login or configure REQUIRE_API_KEY=false for local development.

Endpoints

GET /api/cache

Get cache statistics

bash
curl https://localhost:20128/api/cache \
  -H "Authorization: Bearer $OMNIROUTE_TOKEN"
DELETE /api/cache

Clear all caches

bash
curl -X DELETE https://localhost:20128/api/cache \
  -H "Authorization: Bearer $OMNIROUTE_TOKEN"
GET /api/cache/stats

Get detailed cache statistics

Returns detailed statistics for all cache layers.

bash
curl https://localhost:20128/api/cache/stats \
  -H "Authorization: Bearer $OMNIROUTE_TOKEN"
DELETE /api/cache/stats

Clear cache statistics

bash
curl -X DELETE https://localhost:20128/api/cache/stats \
  -H "Authorization: Bearer $OMNIROUTE_TOKEN"
GET /api/cache/entries

GET cache › entries

bash
curl https://localhost:20128/api/cache/entries \
  -H "Authorization: Bearer $OMNIROUTE_TOKEN"
DELETE /api/cache/entries

DELETE cache › entries

bash
curl -X DELETE https://localhost:20128/api/cache/entries \
  -H "Authorization: Bearer $OMNIROUTE_TOKEN"
GET /api/cache/reasoning

GET cache › reasoning

bash
curl https://localhost:20128/api/cache/reasoning \
  -H "Authorization: Bearer $OMNIROUTE_TOKEN"
DELETE /api/cache/reasoning

DELETE cache › reasoning

bash
curl -X DELETE https://localhost:20128/api/cache/reasoning \
  -H "Authorization: Bearer $OMNIROUTE_TOKEN"

Payloads

See the full OpenAPI specification at GET /api/openapi/spec or docs/openapi.yaml for detailed request/response schemas.

© diegosouzapw, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

Just SKILL.md in skills/omni-cache of diegosouzapw/OmniRoute.

Open the folder on GitHubat commit 8ad6b1c

Used in 1 other repository

We found 1 copy of this SKILL.md (exact, near-identical or edited) in other folders, from 1 other GitHub owner. This page covers the copy in diegosouzapw/OmniRoute, which our catalogue first saw on October 7, 2026.

Compare with similar skills

OmniRoute LLM Cache next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

OmniRoute LLM Cache compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
OmniRoute LLM Cache this skilldiegosouzapw/OmniRoute74k1 repos~529Automated safety check: PassMIT
Caching Architecturemajiayu000/litellm-rs116—~2kAutomated safety check: PassMIT
Vertex AI API DevJetBrains/skills3631 repos~2.4kAutomated safety check: PassNone
Gemini APIgoogle/skills21k3 repos~2.6kAutomated safety check: PassApache-2.0
Gemini Live API Devgoogle-gemini/gemini-skills4.3k—~4.6kAutomated safety check: PassApache-2.0
Cache Credits AnalyzerTsinHzl/kiro2cc-proxy163—~1.1kAutomated safety check: PassMIT

Similar skills

  • Caching Architecture

    majiayu000/litellm-rs

    LiteLLM-RS response caching architecture. An agent skill from majiayu000/litellm-rs.

    116 GitHub stars~2k tokensUpdated today
    AI & LLM EngineeringAuto-check passed
  • Vertex AI API Dev

    JetBrains/skills

    Official

    Guides the usage of Gemini API on Google Cloud Vertex AI with the Gen AI SDK.

    363 GitHub starsUsed in 1 repo~2.4k tokens
    AI & LLM EngineeringAuto-check passed
  • Gemini API

    google/skills

    Official

    A skill your agent uses when the user asks about using Gemini in an enterprise environment or explicitly mentions Vertex AI, Google Cloud, or Agent Platform.

    21k GitHub starsUsed in 3 repos~2.6k tokens
    AI & LLM EngineeringAuto-check passed
  • Gemini Live API Dev

    google-gemini/gemini-skills

    Official

    A skill your agent uses when building real-time, bidirectional streaming applications with the Gemini Live API, or migrating legacy Live models (2.0/2.5/3.1) to Gemini 3.8 Live.

    4.3k GitHub stars~4.6k tokensUpdated today
    Backend & APIsAuto-check passed
  • Cache Credits Analyzer

    TsinHzl/kiro2cc-proxy

    分析 kiro2cc-proxy 访问日志,计算 Prompt Caching 节省的 credits。只要用户粘贴了含有"输入token 输出token 费用$ credits✓"格式的日志行,并询问节省了多少credits、缓存效率、cost分析等,立即使用此 skill。触发关键词:节省了多少credits、cache节省、分析日志、caching…

    163 GitHub stars~1.1k tokensUpdated today
    Backend & APIsAuto-check passed
  • Prompt Caching

    davila7/claude-code-templates

    Caching strategies for LLM prompts including Anthropic prompt caching, response caching, and CAG (Cache Augmented Generation) Use when: prompt caching, cache prompt, response cache, cag, cache…

    32k GitHub starsUsed in 5 repos~452 tokens
    Backend & APIsAuto-check passed

More from diegosouzapw/OmniRoute

All 50 skills in this repo
  • OmniRoute Backup and Sync CLI

    diegosouzapw/OmniRoute

    Backup and restore OmniRoute data from the CLI. Trigger incremental snapshots, sync to cloud storage, manage backup schedules, and restore from archive files.

    74k GitHub stars~948 tokensUpdated today
    Auto-check passed
  • OmniRoute Settings API

    diegosouzapw/OmniRoute

    Read and update global application settings: system prompts, thinking budget, IP filters, payload rules, combo defaults, and require-login configuration.

    74k GitHub starsUsed in 1 repo~3.6k tokens
    Auto-check passed
  • Quality Scan

    diegosouzapw/OmniRoute

    Runs a scoped, read-only quality scan on a repository candidate and reports exact evidence, failures and frozen debt, without treating a static scan as release acceptance.

    74k GitHub stars~748 tokensUpdated today
    Auto-check passed
  • OmniRoute Database Backups

    diegosouzapw/OmniRoute

    Trigger system backups, restore from backup files, and manage the SQLite database lifecycle. Supports export, import, and incremental snapshot strategies.

    74k GitHub starsUsed in 1 repo~395 tokens
    Auto-check passed
  • OmniRoute Provider Management

    diegosouzapw/OmniRoute

    Manages AI provider connections, API keys, OAuth flows and connection tests through OmniRoute's REST API across its 327-provider catalog.

    74k GitHub stars~2.4k tokensUpdated today
    Auto-check passed
  • OmniRoute RTK Context Filters

    diegosouzapw/OmniRoute

    Controls the RTK filter set and context-handling settings in OmniRoute, with endpoints to try compression on sample text and read back retained output.

    74k GitHub starsUsed in 1 repo~618 tokens
    Auto-check passed

Questions about OmniRoute LLM Cache

What does OmniRoute LLM Cache do?

Documents OmniRoute's cache endpoints for reading cache statistics and clearing entries, statistics or the reasoning cache, with notes on TTL and similarity settings. The endpoints shown work on several cache layers. A GET on /api/cache returns cache statistics and a DELETE clears all caches, /api/cache/stats returns detailed statistics for every layer and can clear them, /api/cache/entries lists or deletes entries, and /api/cache/reasoning does the same for reasoning data.

When should I use OmniRoute LLM Cache?

OmniRoute LLM Cache fits situations like: checking how well the LLM response cache is performing; clearing stale cached responses after changing a prompt or model; wiping the reasoning cache without touching other layers; reviewing which entries the cache currently holds.

How do I install OmniRoute LLM Cache in Claude Code?

Run `npx skills add diegosouzapw/OmniRoute --skill omni-cache -a claude-code`. Or copy the skill folder (skills/omni-cache in diegosouzapw/OmniRoute) into .claude/skills/omni-cache in your project. Claude Code loads it when a task matches its description.

How do I install OmniRoute LLM Cache in Codex?

Run `npx skills add diegosouzapw/OmniRoute --skill omni-cache -a codex`. Or copy the skill folder (skills/omni-cache in diegosouzapw/OmniRoute) into .agents/skills/omni-cache in your project. Codex loads it when a task matches its description.

Can I use OmniRoute LLM Cache in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add diegosouzapw/OmniRoute --skill omni-cache -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/omni-cache, .gemini/skills/omni-cache, .github/skills/omni-cache and .opencode/skills/omni-cache in your project.

What does OmniRoute LLM Cache need to run?

Going by SKILL.md and its folder, OmniRoute LLM Cache needs the command-line tools its instructions call (curl) and credentials named OMNIROUTE_TOKEN and REQUIRE_API_KEY. Our summary lists: A running OmniRoute server and a Bearer token or session cookie.

Does OmniRoute LLM Cache access the network?

SKILL.md contains no URLs. Its commands use curl, which can reach the network depending on how they are called. This is read from the text; nothing was executed.

Is OmniRoute LLM Cache safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does OmniRoute LLM Cache use?

OmniRoute LLM Cache is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does OmniRoute LLM Cache use?

About 529 tokens (SKILL.md is roughly 2.1k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to OmniRoute LLM Cache?

Skills that share tags, products or a category with OmniRoute LLM Cache: Caching Architecture (majiayu000/litellm-rs, 116 stars), Vertex AI API Dev (JetBrains/skills, 363 stars), Gemini API (google/skills, 21k stars) and Gemini Live API Dev (google-gemini/gemini-skills, 4.3k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains OmniRoute LLM Cache?

diegosouzapw (a GitHub user) maintains it in diegosouzapw/OmniRoute, which has 73,701 GitHub stars. The repository holds 50 skills in this directory. The repository was last updated on October 6, 2026.

Source: diegosouzapw/OmniRoute on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.