Search
AI & LLM Engineering · llama.cpp · For developers
Skills
Sort:BestMost starsTrending todayTrending this weekTrending this monthNewestRecently updatedName
| # | Skill | Repository | Stars | Used in | Tokens | Auto-check | Licence | Updated |
|---|---|---|---|---|---|---|---|---|
| 1 | Delegate a coding task to Aider (aider) as a background implementer, then review its diff and land it yourself. | amElnagdy/ | 2.3k | 2 repos | ~3k | Automated safety check: Pass | MIT | 4 days ago |
| 2 | Bump or upgrade the pinned versions of Helmor's bundled agent CLIs, SDKs, and supporting binaries — Claude Code + claude-agent-sdk (lockstep), Codex, Cursor SDK, OpenCode, Kimi, Pi, and gh / glab /… | dohooo/ | 1.3k | — | ~2.1k | Automated safety check: Pass | Apache-2.0 | 1 mo ago |
| 3 | Complete agent-ready workflow for Qwen-family MTP or nextn GGUF conversion and release. | R6410418/ | 1.7k | — | ~1.7k | Automated safety check: Pass | MIT | 3 mo ago |
| 4 | Work on vLLM-Omni quantization for diffusion, autoregressive, omni, or multi-stage models. | vllm-project/ | 7.1k | — | ~1.4k | Automated safety check: Pass | Apache-2.0 | today |
| 5 | Trains or fine-tunes language and vision models with TRL or Unsloth on Hugging Face Jobs cloud GPUs, then converts the results to GGUF. | huggingface/ | 11k | 1 repo | ~7.2k | Automated safety check: Pass | Apache-2.0 | 3 days ago |
| 6 | Diagnose OpenAI-compatible model-serving failures from symptoms, endpoint reports, explicit configuration files, or logs while preserving evidence status and requiring confirm/refute checks. | Blackwellboy/ | 135 | — | ~2.1k | Automated safety check: Pass | MIT | 3 days ago |
| 7 | Finds llama.cpp-compatible GGUF models on the Hugging Face Hub, picks a quantization for your hardware and launches them with llama-cli or llama-server. | huggingface/ | 11k | 3 repos | ~945 | Automated safety check: Pass | Apache-2.0 | 3 days ago |
| 8 | 8.Clawmem ClawMem operational reference for agents at query time — the 3-rule escalation gate, MCP tool routing, the 4 query-optimization levers, pipeline behavior (query vs intentsearch), composite scoring… | yoloshii/ | 210 | — | ~7.5k | Automated safety check: Pass | MIT | yesterday |
| 9 | Inspects and tunes the shared-vs-dedicated memory split on AMD Ryzen APUs with unified memory (UMA) so larger LLMs and image-gen models fit on the iGPU, or so reserved GPU memory is returned to the… | amd/ | 408 | — | ~2.6k | Automated safety check: Pass | MIT | 2 days ago |
| 10 | Add a new model export format to AutoRound (e.g., autoround, autogptq, autoawq, gguf, llmcompressor). | intel/ | 1.6k | — | ~1.9k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 11 | 11.Add Model Adapt and port new LLM model architectures to this xinfer project. | guoqingbao/ | 334 | — | ~4.2k | Automated safety check: Notes | MIT | 1 mo ago |
| 12 | Train or fine-tune language models with TRL or Unsloth on Hugging Face Jobs, including SFT, DPO, GRPO, reward models, and GGUF conversion. | waybarrios/ | 534 | — | ~3k | Automated safety check: Pass | Apache-2.0 | 5 days ago |
| 13 | 13.Resolve Always resolve Hugging Face models via model-shelf before any download. | alexziskind1/ | 130 | — | ~792 | Automated safety check: Pass | MIT | 1 mo ago |
| 14 | 14.Test Model Test a model end-to-end using the xybrid execution system. An agent skill from xybrid-ai/xybrid. | xybrid-ai/ | 469 | — | ~1.3k | Automated safety check: Pass | Apache-2.0 | 2 days ago |
| 15 | Uses the Outlines library to constrain model output to a JSON schema, Pydantic model, regex or fixed set of choices when running local models. | Orchestra-Research/ | 13k | 9 repos | ~4k | Automated safety check: Pass | MIT | 3 mo ago |
| 16 | Repository-level wrapper for the canonical Qwen MTP or nextn GGUF release workflow. | R6410418/ | 1.7k | — | ~474 | Automated safety check: Pass | Apache-2.0 | 3 mo ago |
| 17 | 17.Xybrid Init Generate model metadata for an ML model so it works with xybrid. | xybrid-ai/ | 469 | — | ~3k | Automated safety check: Pass | Apache-2.0 | 2 days ago |
| 18 | 18.Test Model Test LLM models served by xinfer for correctness, output quality, and performance. | guoqingbao/ | 334 | — | ~2.6k | Automated safety check: Pass | MIT | 1 mo ago |
| 19 | Maps the /ai/ subsystem of flowfile_core, its three agent tiers, litellm seam, BYOK keys and rate limits, and sets rules for extending or debugging it safely. | Edwardvaneechoud/ | 389 | — | ~9.1k | Automated safety check: Notes | MIT | yesterday |
| 20 | 20.Test Runner Runs the narrowest relevant tests to validate changes. An agent skill from noumena-labs/Sipp. | noumena-labs/ | 121 | — | ~854 | Automated safety check: Pass | Apache-2.0 | 22 days ago |
| 21 | Guided workflow for adding a new model architecture to llama.cpp. | JakeATX/ | 166 | — | ~4.1k | Automated safety check: Pass | MIT | yesterday |
| 22 | 22.Turbofit Hardware-aware adaptive Hermes runtime using portable Turbofiles, total usable memory, owned native llama.cpp residency, stable auto/active:main/active:aux routes, and evidence-backed promotion. | SouthpawIN/ | 107 | — | ~1.9k | Automated safety check: Pass | MIT | 3 days ago |
| 23 | Constrains language model output with regex, selections and grammars using the Guidance library, so JSON, XML, code or formatted fields come out valid. | Orchestra-Research/ | 13k | 5 repos | ~3.6k | Automated safety check: Pass | MIT | 3 mo ago |
| 24 | GGUF format and llama.cpp quantization for efficient CPU/GPU inference. | Orchestra-Research/ | 13k | 3 repos | ~2.6k | Automated safety check: Pass | MIT | 3 mo ago |
| 25 | 25.Llama Cpp Runs LLM inference on CPU, Apple Silicon, and consumer GPUs without NVIDIA hardware. | Orchestra-Research/ | 13k | 3 repos | ~1.5k | Automated safety check: Pass | MIT | 3 mo ago |
| 26 | Hugging Face CLI to estimate the required memory to load Safetensors or GGUF model weights for inference from the Hugging Face Hub | huggingface/ | 11k | 4 repos | ~832 | Automated safety check: Pass | Apache-2.0 | 3 days ago |
| 27 | Register, list, get, and manage LLM models in OCI AI Quick Actions (AQUA) using the ADS SDK. | oracle/ | 125 | — | ~1.4k | Automated safety check: Pass | UPL-1.0 | 1 mo ago |
| 28 | 28.Code Review Review llama.cpp changes against project conventions and common reviewer pitfalls before a PR. | JakeATX/ | 166 | — | ~5.6k | Automated safety check: Pass | MIT | yesterday |
| 29 | A skill your agent uses when converting Hugging Face SafeTensors checkpoints into split BF16 GGUF model repos with skippy-quantize on Hugging Face Jobs or a local machine, then publishing the… | Mesh-LLM/ | 3.5k | — | ~1.1k | Automated safety check: Pass | Apache-2.0 | today |
| 30 | 30.Add Qmd Add QMD (Query Markup Documents) as an advanced memory search backend. | sbusso/ | 194 | — | ~629 | Automated safety check: Pass | MIT | 2 mo ago |
| 31 | Best practices for the Common utilities package in LlamaFarm. | llama-farm/ | 836 | — | ~885 | Automated safety check: Pass | Apache-2.0 | 4 mo ago |
| 32 | Optimize Ollama configuration for the current machine's hardware. | luongnv89/ | 131 | — | ~4.1k | Automated safety check: Notes | MIT | yesterday |
| 33 | A skill your agent uses when changing mesh-llm automation or CLI flows that discover Hugging Face GGUF models, plan CPU Hugging Face Jobs for layer-package splitting, estimate max cost, or publish… | Mesh-LLM/ | 3.5k | — | ~1.4k | Automated safety check: Pass | Apache-2.0 | today |
| 34 | 34.Prime Agent A skill your agent uses when learning, configuring, or troubleshooting Prime Agent (PrimeIntellect-ai/prime-agent), including installation, providers, custom OpenAI-compatible models, local… | wcygan/ | 194 | — | ~1.8k | Automated safety check: Pass | No licence | yesterday |
| 35 | A skill your agent uses when running quantization of a BF16/FP16 GGUF repo and Skippy layer-package creation as one local or Hugging Face Jobs workflow, publishing both artifacts to Hugging Face. | Mesh-LLM/ | 3.5k | — | ~1.6k | Automated safety check: Pass | Apache-2.0 | today |
| 36 | 36.Llama Cpp llama.cpp local GGUF inference + HF Hub model discovery. An agent skill from Tommy-yw/RunbookHermes. | Tommy-yw/ | 546 | 4 repos | ~2.2k | Automated safety check: Pass | MIT | 4 mo ago |
| 37 | 37.App Opinionated app components building on top of ./ui primitives | JakeATX/ | 166 | — | ~146 | Automated safety check: Pass | MIT | yesterday |
| 38 | Train or fine-tune TRL language models on Hugging Face Jobs, including SFT, DPO, GRPO, and GGUF export. | henryalouf/ | 157 | — | ~6.9k | Automated safety check: Pass | MIT | 4 mo ago |
| 39 | Build private, on-device AI features on iPhone, iPad, and Mac with Foundation Models, Core ML, MLX Swift, or llama.cpp. | dpearson2699/ | 1.2k | — | ~3.4k | Automated safety check: Pass | Unknown | 2 mo ago |
| 40 | 40.Pi Agent Builds with and operates Pi, the minimal terminal coding harness. | K-Dense-AI/ | 48k | 1 repo | ~2.1k | Automated safety check: Pass | MIT | 6 days ago |
| 41 | A skill your agent uses when changing mesh-llm's patched llama.cpp Skippy ABI, runtime hooks, model introspection, tensor filtering, activation-frame execution, GGUF writer surface, upstream pin, or… | Mesh-LLM/ | 3.5k | — | ~2.6k | Automated safety check: Pass | Apache-2.0 | today |
| 42 | A skill your agent uses when benchmarking Skippy exact-prefix cache across model families, comparing Skippy against llama-server, producing README benchmark tables, updating… | Mesh-LLM/ | 3.5k | — | ~764 | Automated safety check: Pass | Apache-2.0 | today |
| 43 | Transcribes a single PCM16 WAV file to plain UTF-8 text locally through a standard-library Python wrapper around native SenseVoice Small F16 GGUF FunASR llama.cpp runtimes, preferring cross-vendor… | godot-fun/ | 184 | — | ~707 | Automated safety check: Pass | MIT | today |
| 44 | Train or fine-tune language and vision models using TRL (Transformer Reinforcement Learning) or Unsloth with Hugging Face Jobs infrastructure. | sickn33/ | 47k | 1 repo | ~1.1k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 45 | Master local LLM inference, model selection, VRAM optimization, and local deployment using Ollama, llama.cpp, vLLM, and LM Studio. | sickn33/ | 47k | 2 repos | ~1.6k | Automated safety check: Pass | MIT | 2 days ago |
| 46 | Export a promoted fine-tuned model in the right deployment format — merged safetensors, LoRA-only, GGUF with imatrix, or FP8. | wshobson/ | 40k | — | ~2k | Automated safety check: Pass | MIT | 6 days ago |
| 47 | Find and compare recommended Hugging Face models for a task using benchmarks, model size, and device constraints. | waybarrios/ | 534 | — | ~1.4k | Automated safety check: Pass | Apache-2.0 | 5 days ago |
| 48 | Pick the serving stack and per-runtime memory flags (vLLM, SGLang, llama.cpp, TensorRT Edge-LLM) for an LLM/VLM workload on any NVIDIA Jetson. | NVIDIA/ | 3.6k | 1 repo | ~2.9k | Automated safety check: Pass | Apache-2.0 | 2 days ago |