Agent skill

Local LLM

by coco-research in coco-research/coco

A skill your agent uses when working with this machine's local LLM setup (LM Studio + mlx-dspark) -- checking status, changing context window, diagnosing a reasoning hang or dead request…

Custom licenceAuto-check passedAgent Workflows

Install Local LLM

skills CLI
$ npx skills add coco-research/coco --skill local-llm -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install coco-research/coco local-llm --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/coco-research/coco.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/local-llm .claude/skills/local-llm && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
local-llm
GitHub stars
503
Token cost
~7.9k tokens
SKILL.md length
4,255 words
Files
1
Skills in repo
63
Repo updated
First seen
Licence
Custom licence

At a glance

A skill your agent uses when working with this machine's local LLM setup (LM Studio + mlx-dspark) -- checking status, changing context window, diagnosing a reasoning hang or dead request…

  • Works in 5 steps: kv_bits (KV-cache quantization) --… → lookup-drafts -- currently false. A… → Lower weight quantization below 4-bit --… → …
  • Working with this machines local LLM setup (LM Studio + mlx-dspark) -- checking status
  • SKILL.md covers Announce at start, Architecture, The prefill trick (why this… and Idle-unload (don't keep the…, plus 11 more sections
  • Calls curl; needs MLX_DSPARK_API_KEY

What it does

Local LLM is an agent skill from coco-research/coco. Use when working with this machine's local LLM setup (LM Studio + mlx-dspark) -- checking status, changing context window, diagnosing a reasoning hang or dead request, understanding RAM/speed tradeoffs, or wiring a new script to the local inference endpoint.

Its SKILL.md is about 7.9k tokens, which your agent loads only when the skill is triggered. It is a single SKILL.md file with no bundled scripts.

It sits in Agent Workflows, covering Context engineering. The repository describes itself as: CoCo Super Intelligence is the orchestration layer that turns Claude Code, Cursor, or Codex into an engineering department: a routed advisory board, 226 skills, 386 commands…

When your agent uses it

  • Working with this machines local LLM setup (LM Studio + mlx-dspark) -- checking status
  • Changing context window
  • Diagnosing a reasoning hang
  • Understanding RAM/speed tradeoffs

Example prompts

  • “/local-llm”

Requirements

  • A credential in MLX_DSPARK_API_KEY

Workflow steps

5 steps, taken from the first numbered list in SKILL.md.

  1. kv_bits (KV-cache quantization) -- currently 0/unused. Helps most at
  2. lookup-drafts -- currently false. A cheap draft-generation mode
  3. Lower weight quantization below 4-bit -- diminishing returns, not
  4. max-batch -- already at 4, matching build_local.py's worker pool;
  5. Hardware upgrade (M4 Ultra ~= 2x memory bandwidth) -- out of scope, just

What it can do on your machine

Read from SKILL.md and the folder at commit d79d33a. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Shell commands in SKILL.md call:

    • curl

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md. Its commands use curl, which can reach the network depending on how they are called.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names these keys or tokens, usually read from environment variables:

    • MLX_DSPARK_API_KEY

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Local LLM loads about 7.9k tokens when it runs. Until then it costs about 67 tokens; SKILL.md has 4,255 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~67
When it runs · the whole SKILL.md, loaded when a task matches
~7.9k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

Its licence (Custom licence) doesn't allow us to republish the file, so here is its outline and opening line. It has 4,255 words (~7,943 tokens).

“"I'm using the local-llm skill to work with the local inference setup."”

— opening of SKILL.md by coco-research, Custom licence
name
local-llm
domain
ops

Read the full SKILL.md on GitHub

Files

Just SKILL.md in skills/local-llm of coco-research/coco.

Open the folder on GitHubat commit d79d33a

Compare with similar skills

Local LLM next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Local LLM compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Local LLM this skillcoco-research/coco503—~7.9kAutomated safety check: PassCustom licence
Context Mode Output Sandboxmksglu/context-mode26k—~4.1kAutomated safety check: PassCustom licence
Memori Long-Term MemoryMemoriLabs/Memori17k—~2kAutomated safety check: NotesCustom licence
Picoclaw Skill Creatorsipeed/picoclaw30k—~4.4kAutomated safety check: PassMIT
ccc Semantic Code Searchcocoindex-io/cocoindex-code2.8k—~938Automated safety check: PassApache-2.0
Context Mode for Antigravity CLImksglu/context-mode26k—~850Automated safety check: PassCustom licence

Similar skills

  • Context Mode Output Sandbox

    mksglu/context-mode

    Routes large command, file, API and browser output through context-mode tools so only the needed result enters the agent's context, instead of dumping it via Bash.

    26k GitHub stars~4.1k tokensUpdated today
    Agent WorkflowsAuto-check passed
  • Memori Long-Term Memory

    MemoriLabs/Memori

    Connects Claude Code to Memori Cloud for long-term memory, recalling stored context before substantive replies and saving new context afterward.

    17k GitHub stars~2k tokensUpdated 6 days ago
    Agent WorkflowsAuto-check: notes
  • Picoclaw Skill Creator

    sipeed/picoclaw

    Guidance for creating, updating and reviewing Picoclaw skills, from the SKILL.md structure to organizing bundled scripts, references and assets.

    30k GitHub stars~4.4k tokensUpdated 14 days ago
    Agent WorkflowsAuto-check passed
  • ccc Semantic Code Search

    cocoindex-io/cocoindex-code

    Semantic code search and index management with the ccc CLI: the agent initializes, indexes and queries the project by concept, filtering by language or path.

    2.8k GitHub stars~938 tokensUpdated 2 days ago
    Agent WorkflowsAuto-check passed
  • Routing rules for using context-mode MCP tools in Antigravity CLI: sandboxed code runs, file analysis, indexed search and web fetches that keep large output out of the conversation.

    26k GitHub stars~850 tokensUpdated today
    Agent WorkflowsAuto-check passed
  • Token Optimizer

    alexgreensh/token-optimizer

    Audit a Claude Code or Codex setup for context-window waste, then fix it and measure the savings.

    2.5k GitHub stars~3.6k tokensUpdated today
    Agent WorkflowsAuto-check passed

More from coco-research/coco

All 63 skills in this repo
  • Coco Ads

    coco-research/coco

    Turn the project you just shipped into a short, polished, shareable launch video (an "ad") using HyperFrames.

    503 GitHub stars~1.9k tokensUpdated today
    Auto-check passed
  • Arch Index

    coco-research/coco

    Build and validate .arch/index.json — a committed map from each architectural component of this repository to the real directories and files that implement it, pinned to a git commit, with every…

    503 GitHub stars~2.8k tokensUpdated today
    Auto-check passed
  • Openai Agents

    coco-research/coco

    Build AI applications with OpenAI Agents SDK - text agents, voice agents, multi-agent handoffs, tools with Zod schemas, guardrails, and streaming.

    503 GitHub stars~3.3k tokensUpdated today
    Auto-check passed
  • Skill Evolution

    coco-research/coco

    A skill your agent uses when running, reviewing or changing coco's self-evolution cycle: the 30-day loop that observes how skills are actually used, proposes evidence-backed edits to them as one…

    503 GitHub stars~1.2k tokensUpdated today
    Auto-check passed
  • Coco Diagram

    coco-research/coco

    A skill your agent uses for architecture, current-state, process, data-flow, medallion, or DP diagrams, including redrawing .drawio and Mermaid sources.

    503 GitHub stars~9.2k tokensUpdated today
    Auto-check passed
  • API Design Principles

    coco-research/coco

    A skill your agent uses when designing a new REST or GraphQL API, reviewing an API spec before implementation, setting team API standards, or migrating REST to GraphQL.

    503 GitHub stars~3.4k tokensUpdated today
    Auto-check passed

Categories

Questions about Local LLM

What does Local LLM do?

A skill your agent uses when working with this machine's local LLM setup (LM Studio + mlx-dspark) -- checking status, changing context window, diagnosing a reasoning hang or dead request…. Local LLM is an agent skill from coco-research/coco. Use when working with this machine's local LLM setup (LM Studio + mlx-dspark) -- checking status, changing context window, diagnosing a reasoning hang or dead request, understanding RAM/speed tradeoffs, or wiring a new script to the local inference endpoint.

When should I use Local LLM?

Local LLM fits situations like: working with this machines local LLM setup (LM Studio + mlx-dspark) -- checking status; changing context window; diagnosing a reasoning hang; understanding RAM/speed tradeoffs.

How do I install Local LLM in Claude Code?

Run `npx skills add coco-research/coco --skill local-llm -a claude-code`. Or copy the skill folder (skills/local-llm in coco-research/coco) into .claude/skills/local-llm in your project. Claude Code loads it when a task matches its description.

How do I install Local LLM in Codex?

Run `npx skills add coco-research/coco --skill local-llm -a codex`. Or copy the skill folder (skills/local-llm in coco-research/coco) into .agents/skills/local-llm in your project. Codex loads it when a task matches its description.

Can I use Local LLM in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add coco-research/coco --skill local-llm -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/local-llm, .gemini/skills/local-llm, .github/skills/local-llm and .opencode/skills/local-llm in your project.

What does Local LLM need to run?

Going by SKILL.md and its folder, Local LLM needs the command-line tools its instructions call (curl) and credentials named MLX_DSPARK_API_KEY. Our summary lists: A credential in MLX_DSPARK_API_KEY.

Does Local LLM access the network?

SKILL.md contains no URLs. Its commands use curl, which can reach the network depending on how they are called. This is read from the text; nothing was executed.

Is Local LLM safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Local LLM use?

Local LLM has a licence file (the repository's licence) that doesn't match a standard licence. Read it on GitHub before reusing the skill.

How many tokens does Local LLM use?

About 7.9k tokens (SKILL.md is roughly 32k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Local LLM?

Skills that share tags, products or a category with Local LLM: Context Mode Output Sandbox (mksglu/context-mode, 26k stars), Memori Long-Term Memory (MemoriLabs/Memori, 17k stars), Picoclaw Skill Creator (sipeed/picoclaw, 30k stars) and ccc Semantic Code Search (cocoindex-io/cocoindex-code, 2.8k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Local LLM?

coco-research (a GitHub user) maintains it in coco-research/coco, which has 503 GitHub stars. The repository holds 63 skills in this directory. The repository was last updated on October 9, 2026.

Source: coco-research/coco on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.