AI model or service

llama.cpp agent skills for Claude Code, Codex and other agents.

C/C++ inference engine for running quantised LLMs efficiently on consumer hardware.
skills
89
official
8
Type
AI model or service
Website
github.com
Official GitHub
ggml-org
Reviews
See llama.cpp on Enlisted

llama.cpp skills, ranked

Ranked by score. Sort bymost stars,trending,newest,recently updated

Official (8 skills)

Official llama.cpp skills
#SkillRepositoryStarsUsed inTokensAuto-checkLicenceUpdated
1

Trains or fine-tunes language and vision models with TRL or Unsloth on Hugging Face Jobs cloud GPUs, then converts the results to GGUF.

huggingface/skills11k3 repos~7.2kAutomated safety check: PassApache-2.06 days ago
2

Add a new model export format to AutoRound (e.g., autoround, autogptq, autoawq, gguf, llmcompressor).

intel/auto-round1.6k—~1.9kAutomated safety check: PassApache-2.0today
3

Finds llama.cpp-compatible GGUF models on the Hugging Face Hub, picks a quantization for your hardware and launches them with llama-cli or llama-server.

huggingface/skills11k3 repos~945Automated safety check: PassApache-2.06 days ago
4

Add a new quantization data type to AutoRound (e.g., INT, FP8, MXFP, NVFP, GGUF variants).

intel/auto-round1.6k—~1.5kAutomated safety check: PassApache-2.0today
5
5.Hf MemOfficial

Hugging Face CLI to estimate the required memory to load Safetensors or GGUF model weights for inference from the Hugging Face Hub

huggingface/skills11k4 repos~832Automated safety check: PassApache-2.06 days ago
6

Register, list, get, and manage LLM models in OCI AI Quick Actions (AQUA) using the ADS SDK.

oracle/accelerated-data-science125—~1.4kAutomated safety check: PassUPL-1.01 mo ago
7

Pick the serving stack and per-runtime memory flags (vLLM, SGLang, llama.cpp, TensorRT Edge-LLM) for an LLM/VLM workload on any NVIDIA Jetson.

NVIDIA/skills3.5k1 repo~2.9kAutomated safety check: PassApache-2.0today
8

Benchmark Jetson LLM/VLM serving performance across vLLM, llama.cpp, and Ollama with structured JSON output.

NVIDIA/skills3.5k1 repo~3.1kAutomated safety check: PassApache-2.0today

Community

Community llama.cpp skills
#SkillRepositoryStarsUsed inTokensAuto-checkLicenceUpdated
9

Delegate a coding task to Aider (aider) as a background implementer, then review its diff and land it yourself.

amElnagdy/delegate-skills2.3k3 repos~3kAutomated safety check: PassMITtoday
10

Explains how to configure the Crush coding agent with crushrc or crush.json, covering providers, models, LSPs, MCP servers, hooks, permissions and config precedence.

charmbracelet/crush29k—~3.7kAutomated safety check: PassUnknowntoday
11

Bump or upgrade the pinned versions of Helmor's bundled agent CLIs, SDKs, and supporting binaries — Claude Code + claude-agent-sdk (lockstep), Codex, Cursor SDK, OpenCode, Kimi, Pi, and gh / glab /…

dohooo/helmor1.3k—~2.1kAutomated safety check: PassApache-2.01 mo ago
12

Complete agent-ready workflow for Qwen-family MTP or nextn GGUF conversion and release.

R6410418/Jackrong-llm-finetuning-guide1.7k—~1.7kAutomated safety check: PassMIT2 mo ago
13

Trigger this skill when the user wants to train, fine-tune, or adapt Gemma models (e.g.

google-gemma/gemma-skills1k—~1.9kAutomated safety check: PassApache-2.0yesterday
14

Work on vLLM-Omni quantization for diffusion, autoregressive, omni, or multi-stage models.

vllm-project/vllm-omni7.1k—~1.4kAutomated safety check: PassApache-2.0today
15

Read and write real documents on the device - PDF, XLSX, DOCX, PPTX and CSV.

zhongkaifu/TensorSharp557—~4.2kAutomated safety check: PassBSD-3-Clausetoday
16

Diagnose OpenAI-compatible model-serving failures from symptoms, endpoint reports, explicit configuration files, or logs while preserving evidence status and requiring confirm/refute checks.

Blackwellboy/model-serving-minefield135—~2.1kAutomated safety check: PassMITtoday
17

ClawMem operational reference for agents at query time — the 3-rule escalation gate, MCP tool routing, the 4 query-optimization levers, pipeline behavior (query vs intentsearch), composite scoring…

yoloshii/ClawMem210—~7.5kAutomated safety check: PassMIT2 days ago
18

Inspects and tunes the shared-vs-dedicated memory split on AMD Ryzen APUs with unified memory (UMA) so larger LLMs and image-gen models fit on the iGPU, or so reserved GPU memory is returned to the…

amd/skills398—~2.6kAutomated safety check: PassMITtoday
19

Adapt and port new LLM model architectures to this xinfer project.

guoqingbao/xinfer333—~4.2kAutomated safety check: NotesMIT28 days ago
20

Train or fine-tune language models with TRL or Unsloth on Hugging Face Jobs, including SFT, DPO, GRPO, reward models, and GGUF conversion.

waybarrios/opencode-power-pack533—~3kAutomated safety check: PassApache-2.02 days ago
21

Enforces this monorepo's coding style rules by inspecting git diffs, reading .agents/skills/style-checker/references/styleguidance.md, fixing style violations, and reporting the result.

noumena-labs/Sipp121—~1.4kAutomated safety check: PassApache-2.018 days ago
22

Always resolve Hugging Face models via model-shelf before any download.

alexziskind1/model-shelf130—~792Automated safety check: PassMIT1 mo ago
23

Uses the Outlines library to constrain model output to a JSON schema, Pydantic model, regex or fixed set of choices when running local models.

Orchestra-Research/AI-Research-SKILLs13k10 repos~4kAutomated safety check: PassMIT3 mo ago
24

Test a model end-to-end using the xybrid execution system. An agent skill from xybrid-ai/xybrid.

xybrid-ai/xybrid466—~1.3kAutomated safety check: PassApache-2.0today
25

Use only for current stock/share prices, ticker quotes, and financial market movers (gainers, losers, most-traded shares).

zhongkaifu/TensorSharp557—~1kAutomated safety check: PassBSD-3-Clausetoday
26

Repository-level wrapper for the canonical Qwen MTP or nextn GGUF release workflow.

R6410418/Jackrong-llm-finetuning-guide1.7k—~474Automated safety check: PassApache-2.02 mo ago
27

Generate model metadata for an ML model so it works with xybrid.

xybrid-ai/xybrid466—~3kAutomated safety check: PassApache-2.0today
28

Test LLM models served by xinfer for correctness, output quality, and performance.

guoqingbao/xinfer333—~2.6kAutomated safety check: PassMIT28 days ago
29

A skill your agent uses for web searches and current information lookups, finding sources, fact-checking, researching questions, comparing sources, or summarising web pages.

zhongkaifu/TensorSharp557—~2.3kAutomated safety check: WarnBSD-3-Clausetoday
30

Maps the /ai/ subsystem of flowfile_core, its three agent tiers, litellm seam, BYOK keys and rate limits, and sets rules for extending or debugging it safely.

Edwardvaneechoud/Flowfile373—~7kAutomated safety check: NotesMITtoday
31

GGUF format and llama.cpp quantization for efficient CPU/GPU inference.

Orchestra-Research/AI-Research-SKILLs13k4 repos~2.6kAutomated safety check: PassMIT3 mo ago
32

Runs LLM inference on CPU, Apple Silicon, and consumer GPUs without NVIDIA hardware.

Orchestra-Research/AI-Research-SKILLs13k4 repos~1.5kAutomated safety check: PassMIT3 mo ago
33

Runs the narrowest relevant tests to validate changes. An agent skill from noumena-labs/Sipp.

noumena-labs/Sipp121—~854Automated safety check: PassApache-2.018 days ago
34

Guided workflow for adding a new model architecture to llama.cpp.

JakeATX/llamAmpere148—~4.1kAutomated safety check: PassMITyesterday
35

Hardware-aware adaptive Hermes runtime using portable Turbofiles, total usable memory, owned native llama.cpp residency, stable auto/active:main/active:aux routes, and evidence-backed promotion.

SouthpawIN/turbofit106—~1.9kAutomated safety check: PassMIT3 days ago
36

Constrains language model output with regex, selections and grammars using the Guidance library, so JSON, XML, code or formatted fields come out valid.

Orchestra-Research/AI-Research-SKILLs13k5 repos~3.6kAutomated safety check: PassMIT3 mo ago
37

Review llama.cpp changes against project conventions and common reviewer pitfalls before a PR.

JakeATX/llamAmpere148—~5.2kAutomated safety check: PassMITyesterday
38

A skill your agent uses when converting Hugging Face SafeTensors checkpoints into split BF16 GGUF model repos with skippy-quantize on Hugging Face Jobs or a local machine, then publishing the…

Mesh-LLM/mesh-llm3.5k—~1.1kAutomated safety check: PassApache-2.0today
39

Add QMD (Query Markup Documents) as an advanced memory search backend.

sbusso/claudeclaw194—~629Automated safety check: PassMIT1 mo ago
40

Optimize Ollama configuration for the current machine's hardware.

luongnv89/skills131—~4.1kAutomated safety check: NotesMITtoday
41

A skill your agent uses when creating, monitoring, validating, or documenting low-memory Hugging Face Jobs or local runs that quantize split BF16/FP16 GGUF model repos into custom quant GGUF repos…

Mesh-LLM/mesh-llm3.5k—~1.9kAutomated safety check: PassApache-2.0today
42

A skill your agent uses when changing mesh-llm automation or CLI flows that discover Hugging Face GGUF models, plan CPU Hugging Face Jobs for layer-package splitting, estimate max cost, or publish…

Mesh-LLM/mesh-llm3.5k—~1.4kAutomated safety check: PassApache-2.0today
43

Install and operate Interlinked's optional local semantic function index.

QuentinCody/interlinked-cli178—~1.7kAutomated safety check: PassMIT5 days ago
44

llama.cpp local GGUF inference + HF Hub model discovery. An agent skill from Tommy-yw/RunbookHermes.

Tommy-yw/RunbookHermes5465 repos~2.2kAutomated safety check: PassMIT4 mo ago
45

A skill your agent uses when learning, configuring, or troubleshooting Prime Agent (PrimeIntellect-ai/prime-agent), including installation, providers, custom OpenAI-compatible models, local…

wcygan/dotfiles196—~1.8kAutomated safety check: PassNo licencetoday
46

A skill your agent uses when running quantization of a BF16/FP16 GGUF repo and Skippy layer-package creation as one local or Hugging Face Jobs workflow, publishing both artifacts to Hugging Face.

Mesh-LLM/mesh-llm3.5k—~1.6kAutomated safety check: PassApache-2.0today
47

Best practices for the Common utilities package in LlamaFarm.

llama-farm/llamafarm836—~885Automated safety check: PassApache-2.03 mo ago
48
48.App

Opinionated app components building on top of ./ui primitives

JakeATX/llamAmpere148—~146Automated safety check: PassMITyesterday

Questions, answered from the data.

What is the best llama.cpp skill?

Hugging Face LLM Trainer (official) from huggingface/skills ranks first of the 89 llama.cpp skills listed here, with the highest score: its repository has 11k GitHub stars, 3 other GitHub owners carry a copy, its SKILL.md loads about 7.2k tokens and it passes the automated safety check with no findings. Next come Add Export Format and Hugging Face Local Models.

Is there an official llama.cpp skill?

8 of the 89 llama.cpp skills are official, published by the vendor's own GitHub organization: Hugging Face LLM Trainer, Add Export Format, Hugging Face Local Models, Add Quantization Datatype, Hf Mem and 3 more.

How are these skills ranked?

By Skill Navigator score, which combines the GitHub stars of the skill's repository (shared across that repo's skills and discounted for large collections), how many other GitHub owners carry a copy of the skill, and automated SKILL.md quality checks, minus penalties for safety-check warnings and for each further skill from the same repository. Skills that fail the safety check are not listed.