Search

AI & LLM Engineering · llama.cpp · For developers

70 skills found.
Search results
#SkillRepositoryStarsUsed inTokensAuto-checkLicenceUpdated
1

Delegate a coding task to Aider (aider) as a background implementer, then review its diff and land it yourself.

amElnagdy/delegate-skills2.3k2 repos~3kAutomated safety check: PassMIT4 days ago
2

Bump or upgrade the pinned versions of Helmor's bundled agent CLIs, SDKs, and supporting binaries — Claude Code + claude-agent-sdk (lockstep), Codex, Cursor SDK, OpenCode, Kimi, Pi, and gh / glab /…

dohooo/helmor1.3k—~2.1kAutomated safety check: PassApache-2.01 mo ago
3

Complete agent-ready workflow for Qwen-family MTP or nextn GGUF conversion and release.

R6410418/Jackrong-llm-finetuning-guide1.7k—~1.7kAutomated safety check: PassMIT3 mo ago
4

Work on vLLM-Omni quantization for diffusion, autoregressive, omni, or multi-stage models.

vllm-project/vllm-omni7.1k—~1.4kAutomated safety check: PassApache-2.0today
5

Trains or fine-tunes language and vision models with TRL or Unsloth on Hugging Face Jobs cloud GPUs, then converts the results to GGUF.

huggingface/skills11k1 repo~7.2kAutomated safety check: PassApache-2.03 days ago
6

Diagnose OpenAI-compatible model-serving failures from symptoms, endpoint reports, explicit configuration files, or logs while preserving evidence status and requiring confirm/refute checks.

Blackwellboy/model-serving-minefield135—~2.1kAutomated safety check: PassMIT3 days ago
7

Finds llama.cpp-compatible GGUF models on the Hugging Face Hub, picks a quantization for your hardware and launches them with llama-cli or llama-server.

huggingface/skills11k3 repos~945Automated safety check: PassApache-2.03 days ago
8

ClawMem operational reference for agents at query time — the 3-rule escalation gate, MCP tool routing, the 4 query-optimization levers, pipeline behavior (query vs intentsearch), composite scoring…

yoloshii/ClawMem210—~7.5kAutomated safety check: PassMITyesterday
9

Inspects and tunes the shared-vs-dedicated memory split on AMD Ryzen APUs with unified memory (UMA) so larger LLMs and image-gen models fit on the iGPU, or so reserved GPU memory is returned to the…

amd/skills408—~2.6kAutomated safety check: PassMIT2 days ago
10

Add a new model export format to AutoRound (e.g., autoround, autogptq, autoawq, gguf, llmcompressor).

intel/auto-round1.6k—~1.9kAutomated safety check: PassApache-2.0yesterday
11

Adapt and port new LLM model architectures to this xinfer project.

guoqingbao/xinfer334—~4.2kAutomated safety check: NotesMIT1 mo ago
12

Train or fine-tune language models with TRL or Unsloth on Hugging Face Jobs, including SFT, DPO, GRPO, reward models, and GGUF conversion.

waybarrios/opencode-power-pack534—~3kAutomated safety check: PassApache-2.05 days ago
13

Always resolve Hugging Face models via model-shelf before any download.

alexziskind1/model-shelf130—~792Automated safety check: PassMIT1 mo ago
14

Test a model end-to-end using the xybrid execution system. An agent skill from xybrid-ai/xybrid.

xybrid-ai/xybrid469—~1.3kAutomated safety check: PassApache-2.02 days ago
15

Uses the Outlines library to constrain model output to a JSON schema, Pydantic model, regex or fixed set of choices when running local models.

Orchestra-Research/AI-Research-SKILLs13k9 repos~4kAutomated safety check: PassMIT3 mo ago
16

Repository-level wrapper for the canonical Qwen MTP or nextn GGUF release workflow.

R6410418/Jackrong-llm-finetuning-guide1.7k—~474Automated safety check: PassApache-2.03 mo ago
17

Generate model metadata for an ML model so it works with xybrid.

xybrid-ai/xybrid469—~3kAutomated safety check: PassApache-2.02 days ago
18

Test LLM models served by xinfer for correctness, output quality, and performance.

guoqingbao/xinfer334—~2.6kAutomated safety check: PassMIT1 mo ago
19

Maps the /ai/ subsystem of flowfile_core, its three agent tiers, litellm seam, BYOK keys and rate limits, and sets rules for extending or debugging it safely.

Edwardvaneechoud/Flowfile389—~9.1kAutomated safety check: NotesMITyesterday
20

Runs the narrowest relevant tests to validate changes. An agent skill from noumena-labs/Sipp.

noumena-labs/Sipp121—~854Automated safety check: PassApache-2.022 days ago
21

Guided workflow for adding a new model architecture to llama.cpp.

JakeATX/llamAmpere166—~4.1kAutomated safety check: PassMITyesterday
22

Hardware-aware adaptive Hermes runtime using portable Turbofiles, total usable memory, owned native llama.cpp residency, stable auto/active:main/active:aux routes, and evidence-backed promotion.

SouthpawIN/turbofit107—~1.9kAutomated safety check: PassMIT3 days ago
23

Constrains language model output with regex, selections and grammars using the Guidance library, so JSON, XML, code or formatted fields come out valid.

Orchestra-Research/AI-Research-SKILLs13k5 repos~3.6kAutomated safety check: PassMIT3 mo ago
24

GGUF format and llama.cpp quantization for efficient CPU/GPU inference.

Orchestra-Research/AI-Research-SKILLs13k3 repos~2.6kAutomated safety check: PassMIT3 mo ago
25

Runs LLM inference on CPU, Apple Silicon, and consumer GPUs without NVIDIA hardware.

Orchestra-Research/AI-Research-SKILLs13k3 repos~1.5kAutomated safety check: PassMIT3 mo ago
26
26.Hf MemOfficial

Hugging Face CLI to estimate the required memory to load Safetensors or GGUF model weights for inference from the Hugging Face Hub

huggingface/skills11k4 repos~832Automated safety check: PassApache-2.03 days ago
27

Register, list, get, and manage LLM models in OCI AI Quick Actions (AQUA) using the ADS SDK.

oracle/accelerated-data-science125—~1.4kAutomated safety check: PassUPL-1.01 mo ago
28

Review llama.cpp changes against project conventions and common reviewer pitfalls before a PR.

JakeATX/llamAmpere166—~5.6kAutomated safety check: PassMITyesterday
29

A skill your agent uses when converting Hugging Face SafeTensors checkpoints into split BF16 GGUF model repos with skippy-quantize on Hugging Face Jobs or a local machine, then publishing the…

Mesh-LLM/mesh-llm3.5k—~1.1kAutomated safety check: PassApache-2.0today
30

Add QMD (Query Markup Documents) as an advanced memory search backend.

sbusso/claudeclaw194—~629Automated safety check: PassMIT2 mo ago
31

Best practices for the Common utilities package in LlamaFarm.

llama-farm/llamafarm836—~885Automated safety check: PassApache-2.04 mo ago
32

Optimize Ollama configuration for the current machine's hardware.

luongnv89/skills131—~4.1kAutomated safety check: NotesMITyesterday
33

A skill your agent uses when changing mesh-llm automation or CLI flows that discover Hugging Face GGUF models, plan CPU Hugging Face Jobs for layer-package splitting, estimate max cost, or publish…

Mesh-LLM/mesh-llm3.5k—~1.4kAutomated safety check: PassApache-2.0today
34

A skill your agent uses when learning, configuring, or troubleshooting Prime Agent (PrimeIntellect-ai/prime-agent), including installation, providers, custom OpenAI-compatible models, local…

wcygan/dotfiles194—~1.8kAutomated safety check: PassNo licenceyesterday
35

A skill your agent uses when running quantization of a BF16/FP16 GGUF repo and Skippy layer-package creation as one local or Hugging Face Jobs workflow, publishing both artifacts to Hugging Face.

Mesh-LLM/mesh-llm3.5k—~1.6kAutomated safety check: PassApache-2.0today
36

llama.cpp local GGUF inference + HF Hub model discovery. An agent skill from Tommy-yw/RunbookHermes.

Tommy-yw/RunbookHermes5464 repos~2.2kAutomated safety check: PassMIT4 mo ago
37
37.App

Opinionated app components building on top of ./ui primitives

JakeATX/llamAmpere166—~146Automated safety check: PassMITyesterday
38

Train or fine-tune TRL language models on Hugging Face Jobs, including SFT, DPO, GRPO, and GGUF export.

henryalouf/ruflow157—~6.9kAutomated safety check: PassMIT4 mo ago
39

Build private, on-device AI features on iPhone, iPad, and Mac with Foundation Models, Core ML, MLX Swift, or llama.cpp.

dpearson2699/swift-ios-skills1.2k—~3.4kAutomated safety check: PassUnknown2 mo ago
40

Builds with and operates Pi, the minimal terminal coding harness.

K-Dense-AI/scientific-agent-skills48k1 repo~2.1kAutomated safety check: PassMIT6 days ago
41

A skill your agent uses when changing mesh-llm's patched llama.cpp Skippy ABI, runtime hooks, model introspection, tensor filtering, activation-frame execution, GGUF writer surface, upstream pin, or…

Mesh-LLM/mesh-llm3.5k—~2.6kAutomated safety check: PassApache-2.0today
42

A skill your agent uses when benchmarking Skippy exact-prefix cache across model families, comparing Skippy against llama-server, producing README benchmark tables, updating…

Mesh-LLM/mesh-llm3.5k—~764Automated safety check: PassApache-2.0today
43

Transcribes a single PCM16 WAV file to plain UTF-8 text locally through a standard-library Python wrapper around native SenseVoice Small F16 GGUF FunASR llama.cpp runtimes, preferring cross-vendor…

godot-fun/gai184—~707Automated safety check: PassMITtoday
44

Train or fine-tune language and vision models using TRL (Transformer Reinforcement Learning) or Unsloth with Hugging Face Jobs infrastructure.

sickn33/agentic-awesome-skills47k1 repo~1.1kAutomated safety check: PassApache-2.0yesterday
45

Master local LLM inference, model selection, VRAM optimization, and local deployment using Ollama, llama.cpp, vLLM, and LM Studio.

sickn33/agentic-awesome-skills47k2 repos~1.6kAutomated safety check: PassMIT2 days ago
46

Export a promoted fine-tuned model in the right deployment format — merged safetensors, LoRA-only, GGUF with imatrix, or FP8.

wshobson/agents40k—~2kAutomated safety check: PassMIT6 days ago
47

Find and compare recommended Hugging Face models for a task using benchmarks, model size, and device constraints.

waybarrios/opencode-power-pack534—~1.4kAutomated safety check: PassApache-2.05 days ago
48

Pick the serving stack and per-runtime memory flags (vLLM, SGLang, llama.cpp, TensorRT Edge-LLM) for an LLM/VLM workload on any NVIDIA Jetson.

NVIDIA/skills3.6k1 repo~2.9kAutomated safety check: PassApache-2.02 days ago