Agent skill

Kv Tool Loop Stability

by Mesh-LLM in Mesh-LLM/mesh-llm

A skill your agent uses when certifying mesh-llm KV/cache stability under repeated OpenAI tool-call loops, same-prefix cache reuse, suffix-prefill limits, or native Skippy slot/decode/eviction…

Apache-2.0Auto-check passedAI & LLM Engineering

Install Kv Tool Loop Stability

skills CLI
$ npx skills add Mesh-LLM/mesh-llm --skill kv-tool-loop-stability -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install Mesh-LLM/mesh-llm kv-tool-loop-stability --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/Mesh-LLM/mesh-llm.git skills-src && mkdir -p .claude/skills && cp -r skills-src/.agents/skills/kv-tool-loop-stability .claude/skills/kv-tool-loop-stability && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
kv-tool-loop-stability
GitHub stars
3.5k
Token cost
~704 tokens
SKILL.md length
209 words
Files
1
Skills in repo
25
Repo updated
First seen
Licence
Apache-2.0

At a glance

A skill your agent uses when certifying mesh-llm KV/cache stability under repeated OpenAI tool-call loops, same-prefix cache reuse, suffix-prefill limits, or native Skippy slot/decode/eviction…

  • Works in 5 steps: Attach to an existing OpenAI-compatible… → Prefer a direct model when reproducing… → Run --print-plan first and confirm the… → …
  • Certifying mesh-llm KV/cache stability under repeated OpenAI tool-call loops
  • SKILL.md covers Workflow, Commands, Reporting Rules and Validation
  • Calls python3

What it does

Kv Tool Loop Stability is an agent skill from Mesh-LLM/mesh-llm. Use this skill when certifying mesh-llm KV/cache stability under repeated OpenAI tool-call loops, same-prefix cache reuse, suffix-prefill limits, or native Skippy slot/decode/eviction failures.

Its SKILL.md is about 700 tokens, which your agent loads only when the skill is triggered. It is a single SKILL.md file with no bundled scripts.

It sits in AI & LLM Engineering. It works with OpenAI. The repository describes itself as: Distributed AI/LLM for the people. Share compute privately or publicly to power your agents and chat. The licence is Apache-2.0.

When your agent uses it

  • Certifying mesh-llm KV/cache stability under repeated OpenAI tool-call loops
  • Same-prefix cache reuse
  • Suffix-prefill limits
  • Native Skippy slot/decode/eviction failures

Example prompts

  • “/kv-tool-loop-stability”

Requirements

  • Python 3

Workflow steps

5 steps, taken from the first numbered list in SKILL.md.

  1. Attach to an existing OpenAI-compatible /v1 endpoint. This harness does
  2. Prefer a direct model when reproducing Skippy KV/cache issues. Use auto
  3. Run --print-plan first and confirm the models, attempts,
  4. Pass the active Skippy native log when available. The harness checkpoints
  5. Preserve the evidence directory: manifest.json, results.jsonl,

What it can do on your machine

Read from SKILL.md and the folder at commit 48bf685. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Shell commands in SKILL.md call:

    • python3

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Kv Tool Loop Stability loads about 704 tokens when it runs. Until then it costs about 54 tokens; SKILL.md has 209 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~54
When it runs · the whole SKILL.md, loaded when a task matches
~704

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from Mesh-LLM/mesh-llm at commit 48bf685, republished under its Apache-2.0 licence (© Mesh-LLM). 209 words, ~704 tokens.

Download SKILL.mdSave it as .claude/skills/kv-tool-loop-stability/SKILL.md (or your agent's skills folder).
name
kv-tool-loop-stability
description
Use this skill when certifying mesh-llm KV/cache stability under repeated OpenAI tool-call loops, same-prefix cache reuse, suffix-prefill limits, or native Skippy slot/decode/eviction failures.
metadata.short-description
Certify KV/tool-loop stability

KV Tool-Loop Stability

Use this skill when changing Skippy KV slot cleanup, prefix-cache lookup, OpenAI tool-loop behavior, agent harnesses, or any runtime path related to llama_decode failed, failed to find a memory slot, low same-prefix cache reuse, or proactive eviction failures.

Workflow

  1. Attach to an existing OpenAI-compatible /v1 endpoint. This harness does not start nodes, load models, join meshes, or change routing policy.
  2. Prefer a direct model when reproducing Skippy KV/cache issues. Use auto only when intentionally validating routed behavior.
  3. Run --print-plan first and confirm the models, attempts, pressure_turns, timeout, cache thresholds, output directory, and native logs.
  4. Pass the active Skippy native log when available. The harness checkpoints native logs at run start and scans only appended bytes.
  5. Preserve the evidence directory: manifest.json, results.jsonl, summary.json, summary.md, and transcripts/*.jsonl.

Commands

Preview the run without touching the endpoint:

bash
scripts/qa-kv-tool-loop-stability.py \
  --base-url http://127.0.0.1:9337/v1 \
  --models Qwen/Qwen2.5-3B-Instruct-GGUF:q4_k_m \
  --attempts 5 \
  --pressure-turns 8 \
  --timeout 180 \
  --min-cached-tokens 2048 \
  --suffix-prefill-limit 256 \
  --native-log ~/.mesh-llm/runtime/<pid>/logs/skippy-native.log \
  --output-dir target/kv-tool-loop-stability/local \
  --print-plan

Run the certification:

bash
scripts/qa-kv-tool-loop-stability.py \
  --base-url http://127.0.0.1:9337/v1 \
  --models Qwen/Qwen2.5-3B-Instruct-GGUF:q4_k_m \
  --attempts 5 \
  --pressure-turns 8 \
  --timeout 180 \
  --min-cached-tokens 2048 \
  --suffix-prefill-limit 256 \
  --native-log ~/.mesh-llm/runtime/<pid>/logs/skippy-native.log \
  --output-dir target/kv-tool-loop-stability/local

Reporting Rules

  • Report the model list, attempts, pressure turns, timeout, cache thresholds, success rate, native log paths, and output directory.
  • Include the summary verdict and failing phase details from summary.md or summary.json.
  • Do not paste full prompts, auth headers, huge stable prefixes, or private endpoint data.
  • If no native log is available, say that native-log scanning was not run.

Validation

When changing this harness, run:

bash
python3 -m unittest scripts.tests.test_qa_kv_tool_loop_stability
python3 -m py_compile scripts/qa-kv-tool-loop-stability.py scripts/tests/test_qa_kv_tool_loop_stability.py

© Mesh-LLM, Apache-2.0. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

Just SKILL.md in .agents/skills/kv-tool-loop-stability of Mesh-LLM/mesh-llm.

Open the folder on GitHubat commit 48bf685

Compare with similar skills

Kv Tool Loop Stability next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Kv Tool Loop Stability compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Kv Tool Loop Stability this skillMesh-LLM/mesh-llm3.5k—~704Automated safety check: PassApache-2.0
Chroma Vector DatabaseOrchestra-Research/AI-Research-SKILLs13k8 repos~2.3kAutomated safety check: PassMIT
CLIP Image-Text MatchingOrchestra-Research/AI-Research-SKILLs13k8 repos~1.7kAutomated safety check: PassMIT
Codebase Managementgiancarloerra/SocratiCode3.3k1 repos~1.8kAutomated safety check: PassAGPL-3.0
Azure AI Projects Python SDKmicrosoft/skills3.1k6 repos~2.8kAutomated safety check: PassMIT
Fine-Tuning ExpertJeffallan/claude-skills12k1 repos~1.7kAutomated safety check: PassMIT

Similar skills

  • Chroma Vector Database

    Orchestra-Research/AI-Research-SKILLs

    Shows how to store documents and embeddings in Chroma, query them by similarity with metadata filters, and persist them to disk for RAG and semantic search projects.

    13k GitHub starsUsed in 8 repos~2.3k tokens
    AI & LLM EngineeringAuto-check passed
  • CLIP Image-Text Matching

    Orchestra-Research/AI-Research-SKILLs

    Explains OpenAI's CLIP model for zero-shot image classification, image-text similarity, semantic image search and content moderation, with install steps and code patterns.

    13k GitHub starsUsed in 8 repos~1.7k tokens
    AI & LLM EngineeringAuto-check passed
  • Codebase Management

    giancarloerra/SocratiCode

    Set up, index, and manage SocratiCode codebase indexing. An agent skill from giancarloerra/SocratiCode.

    3.3k GitHub starsUsed in 1 repo~1.8k tokens
    AI & LLM EngineeringAuto-check passed
  • Official

    Reference for building on Microsoft Foundry with the azure-ai-projects Python SDK: project clients, versioned agents, evaluations, connections, datasets and indexes.

    3.1k GitHub starsUsed in 6 repos~2.8k tokens
    AI & LLM EngineeringAuto-check passed
  • Fine-Tuning Expert

    Jeffallan/claude-skills

    Guides LLM fine-tuning with LoRA and QLoRA through Hugging Face PEFT, from dataset validation and training checks to adapter merging, quantization and deployment.

    12k GitHub starsUsed in 1 repo~1.7k tokens
    AI & LLM EngineeringAuto-check passed
  • Docs Planner

    strands-agents/harness-sdk

    Identify documentation gaps and prioritize the docs backlog.

    8.7k GitHub stars~821 tokensUpdated today
    AI & LLM EngineeringAuto-check passed

More from Mesh-LLM/mesh-llm

All 25 skills in this repo
  • Release Validation

    Mesh-LLM/mesh-llm

    A skill your agent uses when validating a MeshLLM release candidate or current HEAD against the last GitHub release, assembling the canonical feature/fix/modification inventory, testing locally…

    3.5k GitHub stars~2.6k tokensUpdated today
    Auto-check passed
  • Benchmark Tune

    Mesh-LLM/mesh-llm

    A skill your agent uses when running, debugging, interpreting, or documenting mesh-llm benchmark tune model-serving throughput trials, including choosing…

    3.5k GitHub stars~1.6k tokensUpdated today
    Auto-check passed
  • A skill your agent uses when adding, renaming, removing, validating, or exposing mesh-llm config settings, including built-in settings, plugin config schemas, owner-control apply behavior, CLI…

    3.5k GitHub stars~1.4k tokensUpdated today
    Auto-check passed
  • Connect Agents

    Mesh-LLM/mesh-llm

    A skill your agent uses when connecting agent tools or OpenAI clients to mesh-llm — launching or configuring Goose, Claude Code, OpenCode, Pi, curl, or any OpenAI-compatible client against a local…

    3.5k GitHub stars~1.2k tokensUpdated today
    Auto-check passed
  • A skill your agent uses when converting Hugging Face SafeTensors checkpoints into split BF16 GGUF model repos with skippy-quantize on Hugging Face Jobs or a local machine, then publishing the…

    3.5k GitHub stars~1.1k tokensUpdated today
    Auto-check passed
  • Hf Gguf Quant Jobs

    Mesh-LLM/mesh-llm

    A skill your agent uses when creating, monitoring, validating, or documenting low-memory Hugging Face Jobs or local runs that quantize split BF16/FP16 GGUF model repos into custom quant GGUF repos…

    3.5k GitHub stars~1.9k tokensUpdated today
    Auto-check passed

Works with

Questions about Kv Tool Loop Stability

What does Kv Tool Loop Stability do?

A skill your agent uses when certifying mesh-llm KV/cache stability under repeated OpenAI tool-call loops, same-prefix cache reuse, suffix-prefill limits, or native Skippy slot/decode/eviction…. Kv Tool Loop Stability is an agent skill from Mesh-LLM/mesh-llm. Use this skill when certifying mesh-llm KV/cache stability under repeated OpenAI tool-call loops, same-prefix cache reuse, suffix-prefill limits, or native Skippy slot/decode/eviction failures.

When should I use Kv Tool Loop Stability?

Kv Tool Loop Stability fits situations like: certifying mesh-llm KV/cache stability under repeated OpenAI tool-call loops; same-prefix cache reuse; suffix-prefill limits; native Skippy slot/decode/eviction failures.

How do I install Kv Tool Loop Stability in Claude Code?

Run `npx skills add Mesh-LLM/mesh-llm --skill kv-tool-loop-stability -a claude-code`. Or copy the skill folder (.agents/skills/kv-tool-loop-stability in Mesh-LLM/mesh-llm) into .claude/skills/kv-tool-loop-stability in your project. Claude Code loads it when a task matches its description.

How do I install Kv Tool Loop Stability in Codex?

Run `npx skills add Mesh-LLM/mesh-llm --skill kv-tool-loop-stability -a codex`. Or copy the skill folder (.agents/skills/kv-tool-loop-stability in Mesh-LLM/mesh-llm) into .agents/skills/kv-tool-loop-stability in your project. Codex loads it when a task matches its description.

Can I use Kv Tool Loop Stability in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add Mesh-LLM/mesh-llm --skill kv-tool-loop-stability -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/kv-tool-loop-stability, .gemini/skills/kv-tool-loop-stability, .github/skills/kv-tool-loop-stability and .opencode/skills/kv-tool-loop-stability in your project.

What does Kv Tool Loop Stability need to run?

Going by SKILL.md and its folder, Kv Tool Loop Stability needs the command-line tools its instructions call (python3). Our summary lists: Python 3.

Does Kv Tool Loop Stability access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is Kv Tool Loop Stability safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Kv Tool Loop Stability use?

Kv Tool Loop Stability is published under the Apache-2.0 licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Kv Tool Loop Stability use?

About 704 tokens (SKILL.md is roughly 2.8k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Kv Tool Loop Stability?

Skills that share tags, products or a category with Kv Tool Loop Stability: Chroma Vector Database (Orchestra-Research/AI-Research-SKILLs, 13k stars), CLIP Image-Text Matching (Orchestra-Research/AI-Research-SKILLs, 13k stars), Codebase Management (giancarloerra/SocratiCode, 3.3k stars) and Azure AI Projects Python SDK (microsoft/skills, 3.1k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Kv Tool Loop Stability?

Mesh-LLM (a GitHub organization) maintains it in Mesh-LLM/mesh-llm, which has 3,485 GitHub stars. The repository holds 25 skills in this directory. The repository was last updated on October 8, 2026.

Source: Mesh-LLM/mesh-llm on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.