Search

llama.cpp

90 skills found, page 2.
Search results
#SkillRepositoryStarsUsed inTokensAuto-checkLicenceUpdated
49

Builds with and operates Pi, the minimal terminal coding harness.

K-Dense-AI/scientific-agent-skills48k1 repo~2.1kAutomated safety check: PassMIT4 days ago
50

A skill your agent uses when changing mesh-llm's llama.cpp patch queue, upstream pin, prepare/build scripts, or carried RPC, MoE, and mesh-hook llama.cpp patches.

Mesh-LLM/mesh-llm3.5k—~1.9kAutomated safety check: PassApache-2.0today
51

A skill your agent uses when changing mesh-llm's patched llama.cpp Skippy ABI, runtime hooks, model introspection, tensor filtering, activation-frame execution, GGUF writer surface, upstream pin, or…

Mesh-LLM/mesh-llm3.5k—~2.6kAutomated safety check: PassApache-2.0today
52

A skill your agent uses when benchmarking Skippy exact-prefix cache across model families, comparing Skippy against llama-server, producing README benchmark tables, updating…

Mesh-LLM/mesh-llm3.5k—~764Automated safety check: PassApache-2.0today
53

A skill your agent uses when inspecting GGUF models, planning layer ranges, generating or validating skippy package artifacts, fake packages for direct GGUFs, materialized stage cache behavior, or…

Mesh-LLM/mesh-llm3.5k—~588Automated safety check: PassApache-2.0today
54

Transcribes a single PCM16 WAV file to plain UTF-8 text locally through a standard-library Python wrapper around native SenseVoice Small F16 GGUF FunASR llama.cpp runtimes, preferring cross-vendor…

godot-fun/gai182—~707Automated safety check: PassMITtoday
55

Train or fine-tune language and vision models using TRL (Transformer Reinforcement Learning) or Unsloth with Hugging Face Jobs infrastructure.

sickn33/agentic-awesome-skills47k1 repo~1.1kAutomated safety check: PassApache-2.0today
56

A skill your agent uses when testing or benchmarking target/draft GGUF pairs for speculative decoding compatibility, tokenizer agreement, draft acceptance rate, or staged verification behavior.

Mesh-LLM/mesh-llm3.5k—~260Automated safety check: PassApache-2.0today
57

A skill your agent uses when certifying a GGUF model family for skippy stage-split serving, reviewing capability data, promoting family evidence into topology policy, or updating staged split…

Mesh-LLM/mesh-llm3.5k—~568Automated safety check: PassApache-2.0today
58

Master local LLM inference, model selection, VRAM optimization, and local deployment using Ollama, llama.cpp, vLLM, and LM Studio.

sickn33/agentic-awesome-skills47k2 repos~1.6kAutomated safety check: PassMITtoday
59

Fine-tune and post-train LLMs with Unsloth Core on a single consumer GPU: VRAM sizing, LoRA/QLoRA, GRPO/DPO, chat-template correctness, and GGUF export.

sickn33/agentic-awesome-skills47k1 repo~4.1kAutomated safety check: PassApache-2.0today
60

Export a promoted fine-tuned model in the right deployment format — merged safetensors, LoRA-only, GGUF with imatrix, or FP8.

wshobson/agents40k—~2kAutomated safety check: PassMIT4 days ago
61

Find and compare recommended Hugging Face models for a task using benchmarks, model size, and device constraints.

waybarrios/opencode-power-pack533—~1.4kAutomated safety check: PassApache-2.03 days ago
62

Pick the serving stack and per-runtime memory flags (vLLM, SGLang, llama.cpp, TensorRT Edge-LLM) for an LLM/VLM workload on any NVIDIA Jetson.

NVIDIA/skills3.5k1 repo~2.9kAutomated safety check: PassApache-2.0today
63

Benchmark Jetson LLM/VLM serving performance across vLLM, llama.cpp, and Ollama with structured JSON output.

NVIDIA/skills3.5k1 repo~3.1kAutomated safety check: PassApache-2.0today
64

Cost per million tokens on hardware you own (Ollama, llama.cpp, vLLM, LM Studio) from watts, electricity price, hardware price and measured tokens/second, and the utilisation at which local beats a…

ruvnet/ruflo74k—~336Automated safety check: NotesMITtoday
65

Build WAN 2.2 First-Last-Frame video workflows. An agent skill from artokun/comfyui-mcp.

artokun/comfyui-mcp800—~5.1kAutomated safety check: PassMIT4 days ago
66

Trace a commit to its published npm versions, including transitive SDK resolution with time-aware accuracy

tetherto/qvac683—~1.7kAutomated safety check: PassApache-2.0today
67

Explains how model files, checkpoints, GGUF quantization, and the Model Manager work in HOT-Step CPP.

scragnog/HOT-Step-CPP173—~6.6kAutomated safety check: PassMIT2 days ago
68

[omh] Self-hosted LLM serving on GPUs: choose the serving engine and quantization from decision tables, prepare deployment as an idempotent runbook with observed-only verification, and measure the…

rlaope/oh-my-hermes3.2k—~2.2kAutomated safety check: PassMITtoday
69

Cross-engine decision rubric for self-hosting or recommending an LLM serving stack.

agentsope/SkillAlchemy466—~6.1kAutomated safety check: PassMITtoday
70

Decision SOP for serving LLMs with vLLM. An agent skill from agentsope/SkillAlchemy.

agentsope/SkillAlchemy466—~6.1kAutomated safety check: PassMITtoday
71

Converts one or more images into faithful text descriptions or OCR with the local 1.3B MiniCPM-V 4.6 GGUF model through llama.cpp, automatically preferring an available Vulkan GPU and falling back…

godot-fun/gai182—~813Automated safety check: PassMITtoday
72

Redact, anonymize, sanitize, or remove PII locally with Distil-PII and llama.cpp; keep personal data and secret values out of model context, logs, and chat.

HybridAIOne/hybridclaw158—~1kAutomated safety check: PassMITtoday
73

Serve an MLX or Hugging Face safetensors model already on this Mac with mlx-lm's server and join it to the person's fleet.

autonomous-ai/openharness1.2k—~1.9kAutomated safety check: PassMITtoday
74

Reuse what Ollama already has on this computer: adopt a running Ollama server into the person's fleet, or serve an Ollama-downloaded model without Ollama.

autonomous-ai/openharness1.2k—~2.2kAutomated safety check: NotesMITtoday
75

Run quick, offline, private LLM tasks on local models via llama.cpp, reusing models already downloaded by Ollama.

glebis/claude-skills390—~1.4kAutomated safety check: PassMITyesterday
76

Prepare export and downstream evaluation handoff for a planned or completed Quark PTQ run.

amd/Quark181—~1.5kAutomated safety check: PassMIT11 days ago
77

在本机 Mac 或 Apple Silicon 上部署 Gemma 4 12B。本地安装/升级 llama.cpp,下载 GGUF 量化模型,用 llama-server 暴露 OpenAI-compatible API,或用 Ollama 暴露本地模型服务;按用户需求在默认 Q4KM、64K/128K 长上下文、QAT Q40 @ 256K、左右对比演示之间选择,配置 tmux…

majiayu000/spellbook287—~875Automated safety check: NotesMITtoday
78

Summarises the delta between a tool's latest release and the last summary the user saw.

sammcj/agentic-coding162—~1.8kAutomated safety check: PassApache-2.0today
79

Set up, install, and configure CONFIDE local de-identification — installs Python deps (natasha, scrubadub, phonenumbers, pymorphy2), ensures Ollama + pulls the default qwen2.5:3b model, detects…

glebis/claude-skills390—~1.1kAutomated safety check: PassMITyesterday
80

A skill your agent uses when running open-weight LLMs locally with Ollama — pulling and tagging models, calling the local API, picking a quantization or GGUF, writing Modelfiles, and sizing VRAM and…

ericrisco/rsc-harness174—~2.8kAutomated safety check: PassMITyesterday
81

Operate, configure, secure, and troubleshoot the LiteLLM AI gateway (proxy) and Python SDK: run the proxy (litellm --config), route to 100+ providers through one OpenAI-compatible API, configure…

magnus919/agent-skills116—~4.2kAutomated safety check: NotesMITtoday
82
82.Vllm

Operate, configure, benchmark, and troubleshoot vLLM inference servers: Docker and Kubernetes deployment, quantization-aware model configuration (tensor parallelism, KV cache), OpenAI-compatible API…

magnus919/agent-skills116—~4.1kAutomated safety check: NotesMITtoday
83

Machine-learning research loop for dataset curation, fine-tuning, evaluation, inference deployment, experiment tracking, and model explainability.

AnastasiyaW/codex-claude-code-config154—~794Automated safety check: PassMITtoday
84

Run quantized LLMs locally with llama.cpp — CPU+GPU inference, GGUF format, OpenAI-compatible server, and Python bindings.

AlexAI-MCP/hermes-CCC135—~2.3kAutomated safety check: PassMIT6 mo ago
85

Build Lightricks LTX-2 / LTX-2.3 video workflows covering text-to-video, image-to-video, GGUF and bundled checkpoints, distilled model, camera control LoRAs, synchronized audio, two-stage upscaling…

artokun/comfyui-mcp800—~6.8kAutomated safety check: WarnMIT4 days ago
86

A skill your agent uses when adapting an open-weight model to a target form or behavior — tone, output format, reasoning pattern — via LoRA/QLoRA or full fine-tuning with TRL SFTTrainer, then…

ericrisco/rsc-harness174—~3.8kAutomated safety check: PassMITyesterday
87

A skill your agent uses when fine-tuning an open-weight LLM fast on ONE GPU with low VRAM — Unsloth's fast model loaders with 4-bit QLoRA and the trl trainer, response-only loss masking so the…

ericrisco/rsc-harness174—~3.6kAutomated safety check: PassMITyesterday
88

Operate, configure, benchmark, and troubleshoot llama.cpp across CPU, Metal, CUDA, HIP/ROCm, Vulkan, SYCL, and hybrid or multi-GPU systems.

magnus919/agent-skills116—~2.3kAutomated safety check: PassMITtoday
89

Plan and execute production ML engineering work — model training and fine-tuning (LoRA/QLoRA), evaluation and eval-set design, quantization decisions, inference deployment, lineage, feature parity…

magnus919/agent-skills116—~1.5kAutomated safety check: PassMITtoday
90

使用本地 BiRefNet GGUF 模型完成图片或视频抠图、人物抠图、主体分割和背景移除,并输出透明 PNG、MOV 或 WebM。适用于用户提到图片抠图、照片去背景、人像透明图、视频抠图、透明视频、BiRefNet、JPG/PNG/BMP/WebP 图片,或 MP4/MOV/WebM 视频的场景;无需 Python、PyTorch 或 CUDA。

aiskillstore/marketplace430—~748Automated safety check: PassMITtoday