AI model or service
llama.cpp agent skills for Claude Code, Codex and other agents.
- skills
- 89
- official
- 8
- Type
- AI model or service
- Website
- github.com
- Official GitHub
- ggml-org
- Reviews
- See llama.cpp on Enlisted
llama.cpp skills, ranked
Ranked by score. Sort bymost stars,trending,newest,recently updated
Official (8 skills)
| # | Skill | Repository | Stars | Used in | Tokens | Auto-check | Licence | Updated |
|---|---|---|---|---|---|---|---|---|
| 1 | Trains or fine-tunes language and vision models with TRL or Unsloth on Hugging Face Jobs cloud GPUs, then converts the results to GGUF. | huggingface/ | 11k | 3 repos | ~7.2k | Automated safety check: Pass | Apache-2.0 | 6 days ago |
| 2 | Add a new model export format to AutoRound (e.g., autoround, autogptq, autoawq, gguf, llmcompressor). | intel/ | 1.6k | — | ~1.9k | Automated safety check: Pass | Apache-2.0 | today |
| 3 | Finds llama.cpp-compatible GGUF models on the Hugging Face Hub, picks a quantization for your hardware and launches them with llama-cli or llama-server. | huggingface/ | 11k | 3 repos | ~945 | Automated safety check: Pass | Apache-2.0 | 6 days ago |
| 4 | Add a new quantization data type to AutoRound (e.g., INT, FP8, MXFP, NVFP, GGUF variants). | intel/ | 1.6k | — | ~1.5k | Automated safety check: Pass | Apache-2.0 | today |
| 5 | Hugging Face CLI to estimate the required memory to load Safetensors or GGUF model weights for inference from the Hugging Face Hub | huggingface/ | 11k | 4 repos | ~832 | Automated safety check: Pass | Apache-2.0 | 6 days ago |
| 6 | Register, list, get, and manage LLM models in OCI AI Quick Actions (AQUA) using the ADS SDK. | oracle/ | 125 | — | ~1.4k | Automated safety check: Pass | UPL-1.0 | 1 mo ago |
| 7 | Pick the serving stack and per-runtime memory flags (vLLM, SGLang, llama.cpp, TensorRT Edge-LLM) for an LLM/VLM workload on any NVIDIA Jetson. | NVIDIA/ | 3.5k | 1 repo | ~2.9k | Automated safety check: Pass | Apache-2.0 | today |
| 8 | Benchmark Jetson LLM/VLM serving performance across vLLM, llama.cpp, and Ollama with structured JSON output. | NVIDIA/ | 3.5k | 1 repo | ~3.1k | Automated safety check: Pass | Apache-2.0 | today |
Community
| # | Skill | Repository | Stars | Used in | Tokens | Auto-check | Licence | Updated |
|---|---|---|---|---|---|---|---|---|
| 9 | Delegate a coding task to Aider (aider) as a background implementer, then review its diff and land it yourself. | amElnagdy/ | 2.3k | 3 repos | ~3k | Automated safety check: Pass | MIT | today |
| 10 | Explains how to configure the Crush coding agent with crushrc or crush.json, covering providers, models, LSPs, MCP servers, hooks, permissions and config precedence. | charmbracelet/ | 29k | — | ~3.7k | Automated safety check: Pass | Unknown | today |
| 11 | Bump or upgrade the pinned versions of Helmor's bundled agent CLIs, SDKs, and supporting binaries — Claude Code + claude-agent-sdk (lockstep), Codex, Cursor SDK, OpenCode, Kimi, Pi, and gh / glab /… | dohooo/ | 1.3k | — | ~2.1k | Automated safety check: Pass | Apache-2.0 | 1 mo ago |
| 12 | Complete agent-ready workflow for Qwen-family MTP or nextn GGUF conversion and release. | R6410418/ | 1.7k | — | ~1.7k | Automated safety check: Pass | MIT | 2 mo ago |
| 13 | Trigger this skill when the user wants to train, fine-tune, or adapt Gemma models (e.g. | google-gemma/ | 1k | — | ~1.9k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 14 | 14.Quantization Work on vLLM-Omni quantization for diffusion, autoregressive, omni, or multi-stage models. | vllm-project/ | 7.1k | — | ~1.4k | Automated safety check: Pass | Apache-2.0 | today |
| 15 | 15.Documents Read and write real documents on the device - PDF, XLSX, DOCX, PPTX and CSV. | zhongkaifu/ | 557 | — | ~4.2k | Automated safety check: Pass | BSD-3-Clause | today |
| 16 | Diagnose OpenAI-compatible model-serving failures from symptoms, endpoint reports, explicit configuration files, or logs while preserving evidence status and requiring confirm/refute checks. | Blackwellboy/ | 135 | — | ~2.1k | Automated safety check: Pass | MIT | today |
| 17 | 17.Clawmem ClawMem operational reference for agents at query time — the 3-rule escalation gate, MCP tool routing, the 4 query-optimization levers, pipeline behavior (query vs intentsearch), composite scoring… | yoloshii/ | 210 | — | ~7.5k | Automated safety check: Pass | MIT | 2 days ago |
| 18 | Inspects and tunes the shared-vs-dedicated memory split on AMD Ryzen APUs with unified memory (UMA) so larger LLMs and image-gen models fit on the iGPU, or so reserved GPU memory is returned to the… | amd/ | 398 | — | ~2.6k | Automated safety check: Pass | MIT | today |
| 19 | 19.Add Model Adapt and port new LLM model architectures to this xinfer project. | guoqingbao/ | 333 | — | ~4.2k | Automated safety check: Notes | MIT | 28 days ago |
| 20 | Train or fine-tune language models with TRL or Unsloth on Hugging Face Jobs, including SFT, DPO, GRPO, reward models, and GGUF conversion. | waybarrios/ | 533 | — | ~3k | Automated safety check: Pass | Apache-2.0 | 2 days ago |
| 21 | Enforces this monorepo's coding style rules by inspecting git diffs, reading .agents/skills/style-checker/references/styleguidance.md, fixing style violations, and reporting the result. | noumena-labs/ | 121 | — | ~1.4k | Automated safety check: Pass | Apache-2.0 | 18 days ago |
| 22 | 22.Resolve Always resolve Hugging Face models via model-shelf before any download. | alexziskind1/ | 130 | — | ~792 | Automated safety check: Pass | MIT | 1 mo ago |
| 23 | Uses the Outlines library to constrain model output to a JSON schema, Pydantic model, regex or fixed set of choices when running local models. | Orchestra-Research/ | 13k | 10 repos | ~4k | Automated safety check: Pass | MIT | 3 mo ago |
| 24 | 24.Test Model Test a model end-to-end using the xybrid execution system. An agent skill from xybrid-ai/xybrid. | xybrid-ai/ | 466 | — | ~1.3k | Automated safety check: Pass | Apache-2.0 | today |
| 25 | 25.Market Data Use only for current stock/share prices, ticker quotes, and financial market movers (gainers, losers, most-traded shares). | zhongkaifu/ | 557 | — | ~1k | Automated safety check: Pass | BSD-3-Clause | today |
| 26 | Repository-level wrapper for the canonical Qwen MTP or nextn GGUF release workflow. | R6410418/ | 1.7k | — | ~474 | Automated safety check: Pass | Apache-2.0 | 2 mo ago |
| 27 | 27.Xybrid Init Generate model metadata for an ML model so it works with xybrid. | xybrid-ai/ | 466 | — | ~3k | Automated safety check: Pass | Apache-2.0 | today |
| 28 | 28.Test Model Test LLM models served by xinfer for correctness, output quality, and performance. | guoqingbao/ | 333 | — | ~2.6k | Automated safety check: Pass | MIT | 28 days ago |
| 29 | 29.Research A skill your agent uses for web searches and current information lookups, finding sources, fact-checking, researching questions, comparing sources, or summarising web pages. | zhongkaifu/ | 557 | — | ~2.3k | Automated safety check: Warn | BSD-3-Clause | today |
| 30 | Maps the /ai/ subsystem of flowfile_core, its three agent tiers, litellm seam, BYOK keys and rate limits, and sets rules for extending or debugging it safely. | Edwardvaneechoud/ | 373 | — | ~7k | Automated safety check: Notes | MIT | today |
| 31 | GGUF format and llama.cpp quantization for efficient CPU/GPU inference. | Orchestra-Research/ | 13k | 4 repos | ~2.6k | Automated safety check: Pass | MIT | 3 mo ago |
| 32 | 32.Llama Cpp Runs LLM inference on CPU, Apple Silicon, and consumer GPUs without NVIDIA hardware. | Orchestra-Research/ | 13k | 4 repos | ~1.5k | Automated safety check: Pass | MIT | 3 mo ago |
| 33 | 33.Test Runner Runs the narrowest relevant tests to validate changes. An agent skill from noumena-labs/Sipp. | noumena-labs/ | 121 | — | ~854 | Automated safety check: Pass | Apache-2.0 | 18 days ago |
| 34 | Guided workflow for adding a new model architecture to llama.cpp. | JakeATX/ | 148 | — | ~4.1k | Automated safety check: Pass | MIT | yesterday |
| 35 | 35.Turbofit Hardware-aware adaptive Hermes runtime using portable Turbofiles, total usable memory, owned native llama.cpp residency, stable auto/active:main/active:aux routes, and evidence-backed promotion. | SouthpawIN/ | 106 | — | ~1.9k | Automated safety check: Pass | MIT | 3 days ago |
| 36 | Constrains language model output with regex, selections and grammars using the Guidance library, so JSON, XML, code or formatted fields come out valid. | Orchestra-Research/ | 13k | 5 repos | ~3.6k | Automated safety check: Pass | MIT | 3 mo ago |
| 37 | 37.Code Review Review llama.cpp changes against project conventions and common reviewer pitfalls before a PR. | JakeATX/ | 148 | — | ~5.2k | Automated safety check: Pass | MIT | yesterday |
| 38 | A skill your agent uses when converting Hugging Face SafeTensors checkpoints into split BF16 GGUF model repos with skippy-quantize on Hugging Face Jobs or a local machine, then publishing the… | Mesh-LLM/ | 3.5k | — | ~1.1k | Automated safety check: Pass | Apache-2.0 | today |
| 39 | 39.Add Qmd Add QMD (Query Markup Documents) as an advanced memory search backend. | sbusso/ | 194 | — | ~629 | Automated safety check: Pass | MIT | 1 mo ago |
| 40 | Optimize Ollama configuration for the current machine's hardware. | luongnv89/ | 131 | — | ~4.1k | Automated safety check: Notes | MIT | today |
| 41 | A skill your agent uses when creating, monitoring, validating, or documenting low-memory Hugging Face Jobs or local runs that quantize split BF16/FP16 GGUF model repos into custom quant GGUF repos… | Mesh-LLM/ | 3.5k | — | ~1.9k | Automated safety check: Pass | Apache-2.0 | today |
| 42 | A skill your agent uses when changing mesh-llm automation or CLI flows that discover Hugging Face GGUF models, plan CPU Hugging Face Jobs for layer-package splitting, estimate max cost, or publish… | Mesh-LLM/ | 3.5k | — | ~1.4k | Automated safety check: Pass | Apache-2.0 | today |
| 43 | Install and operate Interlinked's optional local semantic function index. | QuentinCody/ | 178 | — | ~1.7k | Automated safety check: Pass | MIT | 5 days ago |
| 44 | 44.Llama Cpp llama.cpp local GGUF inference + HF Hub model discovery. An agent skill from Tommy-yw/RunbookHermes. | Tommy-yw/ | 546 | 5 repos | ~2.2k | Automated safety check: Pass | MIT | 4 mo ago |
| 45 | 45.Prime Agent A skill your agent uses when learning, configuring, or troubleshooting Prime Agent (PrimeIntellect-ai/prime-agent), including installation, providers, custom OpenAI-compatible models, local… | wcygan/ | 196 | — | ~1.8k | Automated safety check: Pass | No licence | today |
| 46 | A skill your agent uses when running quantization of a BF16/FP16 GGUF repo and Skippy layer-package creation as one local or Hugging Face Jobs workflow, publishing both artifacts to Hugging Face. | Mesh-LLM/ | 3.5k | — | ~1.6k | Automated safety check: Pass | Apache-2.0 | today |
| 47 | Best practices for the Common utilities package in LlamaFarm. | llama-farm/ | 836 | — | ~885 | Automated safety check: Pass | Apache-2.0 | 3 mo ago |
| 48 | 48.App Opinionated app components building on top of ./ui primitives | JakeATX/ | 148 | — | ~146 | Automated safety check: Pass | MIT | yesterday |
Questions, answered from the data.
What is the best llama.cpp skill?
Hugging Face LLM Trainer (official) from huggingface/skills ranks first of the 89 llama.cpp skills listed here, with the highest score: its repository has 11k GitHub stars, 3 other GitHub owners carry a copy, its SKILL.md loads about 7.2k tokens and it passes the automated safety check with no findings. Next come Add Export Format and Hugging Face Local Models.
Is there an official llama.cpp skill?
8 of the 89 llama.cpp skills are official, published by the vendor's own GitHub organization: Hugging Face LLM Trainer, Add Export Format, Hugging Face Local Models, Add Quantization Datatype, Hf Mem and 3 more.
How are these skills ranked?
By Skill Navigator score, which combines the GitHub stars of the skill's repository (shared across that repo's skills and discounted for large collections), how many other GitHub owners carry a copy of the skill, and automated SKILL.md quality checks, minus penalties for safety-check warnings and for each further skill from the same repository. Skills that fail the safety check are not listed.