Search

Qwen · LLM inference and serving

31 skills found.
Search results
#SkillRepositoryStarsUsed inTokensAuto-checkLicenceUpdated
1

Diagnose and optimize vLLM Omni diffusion workloads, especially Wan/Qwen/Flux-style image and video generation.

vllm-project/vllm-omni7.1k—~7.5kAutomated safety check: PassApache-2.0today
2

Complete agent-ready workflow for Qwen-family MTP or nextn GGUF conversion and release.

R6410418/Jackrong-llm-finetuning-guide1.7k—~1.7kAutomated safety check: PassMIT3 mo ago
3

Run wally's LLM e2e on Apple Neural Engine (NeuRT) and Snapdragon Hexagon NPU (QHexRT) devices.

RunanywhereAI/wally1.6k—~845Automated safety check: PassMITtoday
4

Adapt and port new LLM model architectures to this xinfer project.

guoqingbao/xinfer334—~4.2kAutomated safety check: NotesMIT1 mo ago
5

Always resolve Hugging Face models via model-shelf before any download.

alexziskind1/model-shelf130—~792Automated safety check: PassMIT1 mo ago
6

Check model compatibility with xinfer before loading. An agent skill from guoqingbao/xinfer.

guoqingbao/xinfer334—~3.8kAutomated safety check: PassMIT1 mo ago
7

Use only for current stock/share prices, ticker quotes, and financial market movers (gainers, losers, most-traded shares).

zhongkaifu/TensorSharp568—~1kAutomated safety check: PassBSD-3-Clausetoday
8

Test LLM models served by xinfer for correctness, output quality, and performance.

guoqingbao/xinfer334—~2.6kAutomated safety check: PassMIT1 mo ago
9

A skill your agent uses for web searches and current information lookups, finding sources, fact-checking, researching questions, comparing sources, or summarising web pages.

zhongkaifu/TensorSharp568—~2.3kAutomated safety check: WarnBSD-3-Clausetoday
10

Guided workflow for adding a new model architecture to llama.cpp.

JakeATX/llamAmpere166—~4.1kAutomated safety check: PassMITyesterday
11

Hardware-aware adaptive Hermes runtime using portable Turbofiles, total usable memory, owned native llama.cpp residency, stable auto/active:main/active:aux routes, and evidence-backed promotion.

SouthpawIN/turbofit107—~1.9kAutomated safety check: PassMIT3 days ago
12

Serves AI models on AMD Instinct GPU hardware using vLLM. An agent skill from amd/skills.

amd/skills408—~4kAutomated safety check: NotesMIT2 days ago
13

Guide a tester through hipfire bring-up, serve smoke, claim-scoped harnesses, and benchmark reporting on AMD RDNA/CDNA GPUs.

warpfront/hipfire658—~1.5kAutomated safety check: PassUnknownyesterday
14

Provides guidance for enterprise-grade RL training using miles, a production-ready fork of slime.

Orchestra-Research/AI-Research-SKILLs13k2 repos~2.2kAutomated safety check: PassMIT3 mo ago
15

Review llama.cpp changes against project conventions and common reviewer pitfalls before a PR.

JakeATX/llamAmpere166—~5.6kAutomated safety check: PassMITyesterday
16

Breaks LLM torch profiler traces down by forward pass, layer and kernel, with timing tables and Perfetto time ranges for the layers you want to inspect.

BBuf/AI-Infra-Auto-Driven-SKILLS938—~3.9kAutomated safety check: PassNo licence6 days ago
17

A skill your agent uses when a new Ollama Cloud model is announced or available (e.g.

heypinchy/pinchy182—~3.9kAutomated safety check: NotesAGPL-3.020 days ago
18
18.App

Opinionated app components building on top of ./ui primitives

JakeATX/llamAmpere166—~146Automated safety check: PassMITyesterday
19

Embed alibaba/page-agent into your own web application — a pure-JavaScript in-page GUI agent that ships as a single <script tag or npm package and lets end-users of your site drive the UI with…

Tommy-yw/RunbookHermes5463 repos~2.3kAutomated safety check: NotesMIT4 mo ago
20

Deploy and serve local models with Ollama — pull and run them, then expose the OpenAI-compatible endpoint to apps and agents.

Prism-Shadow/penguin-harness2.5k—~839Automated safety check: NotesApache-2.0yesterday
21
21.Vllm

Deploy and serve LLMs with vLLM behind an OpenAI-compatible endpoint, with tool calling enabled for agent workloads.

Prism-Shadow/penguin-harness2.5k—~1kAutomated safety check: PassApache-2.0yesterday
22

Explains how model files, checkpoints, GGUF quantization, and the Model Manager work in HOT-Step CPP.

scragnog/HOT-Step-CPP174—~6.6kAutomated safety check: PassMIT2 days ago
23

Low-memory file2file quantization for very large safetensors LLMs that cannot be loaded whole.

amd/Quark182—~2.3kAutomated safety check: PassMIT13 days ago
24

L3 recipe that runs a Torch LLM PTQ end-to-end for AMD Quark — for PyTorch / HuggingFace transformers models (safetensors input): quantize → validate → evaluate.

amd/Quark182—~2.6kAutomated safety check: PassMIT13 days ago
25

Torch LLM PTQ workflow for AMD Quark — from model selection to quantized output.

amd/Quark182—~2.8kAutomated safety check: PassMIT13 days ago
26

Inspect a target model and prepare metadata for Quark PTQ planning.

amd/Quark182—~2kAutomated safety check: PassMIT13 days ago
27

Plan or review LoRA and edit-training work specifically for FLUX.2 Klein or Qwen-Image-Edit, including paired datasets, trainer-version contracts, and held-out fidelity checks.

AnastasiyaW/codex-claude-code-config154—~4.5kAutomated safety check: PassMITyesterday
28

Runs an end-to-end AMD Quark post-training quantization workflow for PyTorch / Hugging Face LLMs: inspect a Hub or local model, choose a quantization plan, create reproducible artifacts, request…

amd/Quark182—~2.3kAutomated safety check: PassMIT13 days ago
29

vLLM Ascend plugin for LLM inference serving on Huawei Ascend NPU.

ascend-ai-coding/awesome-ascend-skills174—~2.7kAutomated safety check: PassNo licenceyesterday
30

Track daily PRs and Issues from vllm-project/vllm and vllm-project/vllm-ascend, filter by model (DeepSeek/Qwen/GLM/MiniMax/Kimi) and tech topics (PD disaggregation, MTP, quantization, graph mode…

ascend-ai-coding/awesome-ascend-skills174—~731Automated safety check: PassNo licenceyesterday
31

昇腾 NPU 平台 vLLM 大模型推理服务一键部署。触发:用户说'部署 模型名'、'NPU 部署模型'、'vllm serve'。流程:SSH检查 → NPU检查 → 配置发现(必须验证) → 用户确认 → 部署 → cron监控 → 验证。约束:(1) 配置必须从官方文档验证,禁止猜测;(2) 后台启动必须用cron监控,禁止手动轮询。支持…

ascend-ai-coding/awesome-ascend-skills174—~1.2kAutomated safety check: PassNo licenceyesterday