AI model or service

SGLang agent skills for Claude Code, Codex and other agents.

Fast serving framework for large language and vision-language models.
skills
79
official
6
Type
AI model or service
Website
docs.sglang.ai
Official GitHub
sgl-project
Reviews
See SGLang on Enlisted

SGLang skills, ranked

Ranked by score. Sort bymost stars,trending,newest,recently updated

Official (6 skills)

Official SGLang skills
#SkillRepositoryStarsUsed inTokensAuto-checkLicenceUpdated
1

Chooses the right serving container and current image URI for deploying a Hugging Face model to a SageMaker endpoint, preferring Hugging Face images over generic ones.

huggingface/skills11k1 repo~4.6kAutomated safety check: PassApache-2.06 days ago
2
2.Debug InferenceOfficial

Debug inference clients that use an attached provider and its native endpoint, including hosted APIs and host-local Ollama, vLLM, SGLang, TRT-LLM, LM Studio, or NIM.

NVIDIA/OpenShell15k—~1.9kAutomated safety check: PassApache-2.0today
3

Add a new model export format to AutoRound (e.g., autoround, autogptq, autoawq, gguf, llmcompressor).

intel/auto-round1.6k—~1.9kAutomated safety check: PassApache-2.0yesterday
4

Measure Jetson DRAM/NvMap usage and verify before/after memory reclamation with live audit data.

NVIDIA/skills3.5k1 repo~2.3kAutomated safety check: NotesApache-2.0today
5

Stand up vLLM or SGLang serving on Jetson, using upstream vLLM on Thor and Orin JetPack 7.2+, and NVIDIA-AI-IOT vLLM on older Orin.

NVIDIA/skills3.5k2 repos~3kAutomated safety check: NotesApache-2.0today
6

Pick the serving stack and per-runtime memory flags (vLLM, SGLang, llama.cpp, TensorRT Edge-LLM) for an LLM/VLM workload on any NVIDIA Jetson.

NVIDIA/skills3.5k1 repo~2.9kAutomated safety check: PassApache-2.0today

Community

Community SGLang skills
#SkillRepositoryStarsUsed inTokensAuto-checkLicenceUpdated
7

Replay-first debug flow for SGLang serving problems. An agent skill from sgl-project/sglang.

sgl-project/sglang37k3 repos~2.1kAutomated safety check: PassApache-2.0today
8

Unified LLM torch-profiler triage skill for sglang, vllm, TensorRT-LLM, and TokenSpeed.

sgl-project/sglang37k2 repos~6.4kAutomated safety check: PassApache-2.0today
9

Start and persistently pursue a goal to babysit an SGLang pull request until selected GitHub Actions workflows pass on the latest PR head.

sgl-project/sglang37k2 repos~3kAutomated safety check: PassApache-2.0today
10

Debug hanging issues in SGLang distributed inference (TP/PP/DP/EP).

sgl-project/sglang37k2 repos~2.4kAutomated safety check: PassApache-2.0today
11

Conventions for SGLang environment variables — where to define, how to access, how to name, and how to deprecate.

sgl-project/sglang37k2 repos~2.9kAutomated safety check: PassApache-2.0today
12

Write, calibrate, and debug the prefill-vs-decode logprob (KL) consistency tests in sglang -- the two independent conditions a zero requires (every operator batch-invariant, and the two paths…

sgl-project/sglang37k2 repos~3.7kAutomated safety check: PassApache-2.0today
13

Plans and audits Day-0 SGLang support for a new model release: scope, architecture gaps, PR order, validation gates and sanitized public evidence.

BBuf/AI-Infra-Auto-Driven-SKILLS900—~2.3kAutomated safety check: PassNo licence2 days ago
14

Code style for SGLang large classes Scheduler, TokenizerManager, and ModelRunner: frozen-code conventions and init orchestration style.

sgl-project/sglang37k2 repos~1.9kAutomated safety check: PassApache-2.0today
15

Use with the dstack skill for model-serving work when the image, serving command, resources, backend/fleet choice, or service behavior is not proven.

dstackai/dstack2.3k—~1.6kAutomated safety check: PassMPL-2.0today
16

Analyze, benchmark, diagnose, and optimize large-model inference deployments from hardware inventory, model details, workload traces, and latency or throughput SLOs.

rednote-machine-learning/Inference-autopilot142—~4.5kAutomated safety check: PassApache-2.01 mo ago
17

Reads SGLang or vLLM startup logs to show where GPU memory went and estimates how many concurrent requests fit at common token lengths.

BBuf/AI-Infra-Auto-Driven-SKILLS900—~2.5kAutomated safety check: PassNo licence2 days ago
18

Profile AI inference workloads (vLLM, SGLang, TensorRT-LLM, PyTorch, any GPU application) with the Graphsignal profiler and read the results from its local /signals JSON endpoint.

graphsignal/graphsignal257—~6.2kAutomated safety check: PassApache-2.09 days ago
19

Clean up noisy startup warnings and spurious prints in SGLang server logs.

guqiong96/Lsglang1431 repo~4.5kAutomated safety check: PassApache-2.02 days ago
20

Trigger the bot-cherry-pick workflow for a batch of merged PRs onto a release branch and monitor each run to completion.

sgl-project/sglang37k2 repos~4kAutomated safety check: PassApache-2.0today
21

驱动 One-Eval 对 API 或本地模型做端到端评测,覆盖纯文本、多模态、代码生成、函数调用和 Agent benchmark。当用户想评测模型在一个或多个 benchmark 上的表现、比较分数、补充 metric,或生成图文评测报告时使用本 skill。

OpenDCAI/One-Eval165—~2.4kAutomated safety check: PassApache-2.01 mo ago
22

Analyzes Torch Profiler traces from SGLang, vLLM and TensorRT-LLM servers into kernel attribution, overlap and fusion tables.

BBuf/AI-Infra-Auto-Driven-SKILLS900—~2.8kAutomated safety check: PassNo licence2 days ago
23

Diagnose and correct GPT-QModel tokenizer initialization, tokenization normalization, special-token handling, prompt rendering, and chat-template problems.

ModelCloud/GPTQModel1.3k—~1.1kAutomated safety check: PassUnknowntoday
24

A skill your agent uses when benchmarking denoise latency or profiling a diffusion bottleneck in SGLang.

sgl-project/sglang37k2 repos~2.4kAutomated safety check: PassApache-2.0today
25

Looks up public original architecture diagrams for named LLM, vision-language, MoE, diffusion and OCR models and returns the image with its source attribution.

BBuf/AI-Infra-Auto-Driven-SKILLS900—~1.2kAutomated safety check: PassNo licence2 days ago
26

Fallback installer for milesdiffusion on a bare CUDA 12.9 Linux GPU box, reproducing the official radixark/milesdiffusion image's package versions and verifying them.

radixark/miles_diffusion107—~1.6kAutomated safety check: PassApache-2.0today
27

A skill your agent uses when quantizing a diffusion DiT with NVIDIA ModelOpt and making the resulting FP8 or NVFP4 checkpoint loadable, verifiable, and benchmarkable in SGLang Diffusion.

sgl-project/sglang37k2 repos~5kAutomated safety check: PassApache-2.0today
28

Guides reinforcement-learning post-training of LLMs with slime, which pairs Megatron-LM training with SGLang rollouts, including GRPO runs on GLM, Qwen3 and Llama 3 models.

Orchestra-Research/AI-Research-SKILLs13k5 repos~2.8kAutomated safety check: PassMIT3 mo ago
29

Step-by-step tutorial for adding a new lightweight JIT CUDA kernel to sglang's jitkernel module

guqiong96/Lsglang1431 repo~10kAutomated safety check: PassApache-2.02 days ago
30

Naming conventions for SGLang speculative decoding identifiers.

sgl-project/sglang37k2 repos~1.6kAutomated safety check: PassApache-2.0today
31

Validate changed quantization math, packed formats, and inference kernels against a Torch oracle; adjudicate near-tie mismatches before rejecting optimized arithmetic.

ModelCloud/GPTQModel1.3k—~1.4kAutomated safety check: PassUnknowntoday
32

Reviews SGLang changes the way its maintainers do, drawing on a bundled corpus of public PR review threads and a flowchart of how the diff runs.

BBuf/AI-Infra-Auto-Driven-SKILLS900—~4.6kAutomated safety check: PassNo licence2 days ago
33

Sets up a structured debugging session for a Dynamo bug — pull the report from a Linear ticket, GitHub issue, or pasted text, capture the environment, create a persistent worklog markdown file, and…

ai-dynamo/dynamo8.2k—~1.2kAutomated safety check: PassApache-2.0today
34

Add a new model to the SGLang Cookbook (docs/, Mintlify), config-driven format — instantiate the model-agnostic template into a per-model config (+ benchmarks) JSX under src/snippets/configs/, an…

sgl-project/sglang37k2 repos~4kAutomated safety check: PassApache-2.0today
35

Migrate a legacy-template SGLang cookbook page (monolithic per-model generator under docs/src/snippets/autoregressive/) onto the config-driven template (shared deployment.jsx / playground.jsx…

sgl-project/sglang37k2 repos~4kAutomated safety check: PassApache-2.0today
36

Guide to SGLang CI workflow orchestration — stage ordering, fail-fast, gating, partitioning, execution modes, and debugging CI failures.

sgl-project/sglang37k2 repos~5.5kAutomated safety check: PassApache-2.0today
37

A skill your agent uses when choosing the fastest SGLang Diffusion flags for a model, GPU, and VRAM budget.

sgl-project/sglang37k2 repos~13kAutomated safety check: PassApache-2.0today
38

Guide for writing SGLang CI/UT tests. An agent skill from sgl-project/sglang.

sgl-project/sglang37k2 repos~5.2kAutomated safety check: PassApache-2.0today
39

Covers serving LLMs with SGLang, whose RadixAttention reuses cached prefixes, and constraining output to JSON, regex or grammar for agent and tool-calling workloads.

Orchestra-Research/AI-Research-SKILLs13k3 repos~2.9kAutomated safety check: PassMIT3 mo ago
40

Trains LLMs with reinforcement learning using verl, from ByteDance's Seed team, with GRPO, PPO and other algorithms and swappable training and rollout backends.

Orchestra-Research/AI-Research-SKILLs13k3 repos~2.4kAutomated safety check: PassMIT3 mo ago
41

Step-by-step tutorial for adding a heavyweight AOT CUDA/C++ kernel to sgl-kernel (including tests & benchmarks)

sgl-project/sglang37k2 repos~3.4kAutomated safety check: PassApache-2.0today
42

Review a pull request against the SGLang Cookbook (docs/, Mintlify) contribution checklist — the config-driven format (per-model config + benchmarks JSX consumed by the shared deployment.jsx /…

sgl-project/sglang37k2 repos~4.2kAutomated safety check: PassApache-2.0today
43

Call this skill when you need to debug CUDA crashes in SGLang using kernel API logging

sgl-project/sglang37k2 repos~4.9kAutomated safety check: PassApache-2.0today
44

Generate an e2e profiling trace of an SGLang server run. An agent skill from sgl-project/sglang.

sgl-project/sglang37k2 repos~1.1kAutomated safety check: PassApache-2.0today
45

Requirements for the SGLang scripted runtime, chiefly when to add (vs not add) a harness API.

sgl-project/sglang37k2 repos~448Automated safety check: PassApache-2.0today
46

Investigate consistently failing SGLang CI tests by extracting the failure signature from scheduled or rerun workflows, bisecting the passing/failing commit window, checking runner or hardware…

sgl-project/sglang37k2 repos~2.5kAutomated safety check: PassApache-2.0today
47

A skill your agent uses when adding a new diffusion model or Diffusers pipeline to SGLang.

sgl-project/sglang37k2 repos~9.9kAutomated safety check: PassApache-2.0today
48

Run RecIF beam×concurrency eval and FlashRec vs SGLang/vLLM/TRT-LLM baselines.

sohu-mptc/FlashRec107—~915Automated safety check: PassApache-2.05 days ago

Questions, answered from the data.

What is the best SGLang skill?

SageMaker Serving Image Selection (official) from huggingface/skills ranks first of the 79 SGLang skills listed here, with the highest score: its repository has 11k GitHub stars, 1 other GitHub owner carry a copy, its SKILL.md loads about 4.6k tokens and it passes the automated safety check with no findings. Next come Debug Inference and Add Export Format.

Is there an official SGLang skill?

6 of the 79 SGLang skills are official, published by the vendor's own GitHub organization: SageMaker Serving Image Selection, Debug Inference, Add Export Format, Jetson Memory Audit, Jetson LLM Serve and 1 more.

How are these skills ranked?

By Skill Navigator score, which combines the GitHub stars of the skill's repository (shared across that repo's skills and discounted for large collections), how many other GitHub owners carry a copy of the skill, and automated SKILL.md quality checks, minus penalties for safety-check warnings and for each further skill from the same repository. Skills that fail the safety check are not listed.