AI model or service
SGLang agent skills for Claude Code, Codex and other agents.
- skills
- 79
- official
- 6
- Type
- AI model or service
- Website
- docs.sglang.ai
- Official GitHub
- sgl-project
- Reviews
- See SGLang on Enlisted
SGLang skills, ranked
Ranked by score. Sort bymost stars,trending,newest,recently updated
Official (6 skills)
| # | Skill | Repository | Stars | Used in | Tokens | Auto-check | Licence | Updated |
|---|---|---|---|---|---|---|---|---|
| 1 | Chooses the right serving container and current image URI for deploying a Hugging Face model to a SageMaker endpoint, preferring Hugging Face images over generic ones. | huggingface/ | 11k | 1 repo | ~4.6k | Automated safety check: Pass | Apache-2.0 | 6 days ago |
| 2 | Debug inference clients that use an attached provider and its native endpoint, including hosted APIs and host-local Ollama, vLLM, SGLang, TRT-LLM, LM Studio, or NIM. | NVIDIA/ | 15k | — | ~1.9k | Automated safety check: Pass | Apache-2.0 | today |
| 3 | Add a new model export format to AutoRound (e.g., autoround, autogptq, autoawq, gguf, llmcompressor). | intel/ | 1.6k | — | ~1.9k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 4 | Measure Jetson DRAM/NvMap usage and verify before/after memory reclamation with live audit data. | NVIDIA/ | 3.5k | 1 repo | ~2.3k | Automated safety check: Notes | Apache-2.0 | today |
| 5 | Stand up vLLM or SGLang serving on Jetson, using upstream vLLM on Thor and Orin JetPack 7.2+, and NVIDIA-AI-IOT vLLM on older Orin. | NVIDIA/ | 3.5k | 2 repos | ~3k | Automated safety check: Notes | Apache-2.0 | today |
| 6 | Pick the serving stack and per-runtime memory flags (vLLM, SGLang, llama.cpp, TensorRT Edge-LLM) for an LLM/VLM workload on any NVIDIA Jetson. | NVIDIA/ | 3.5k | 1 repo | ~2.9k | Automated safety check: Pass | Apache-2.0 | today |
Community
| # | Skill | Repository | Stars | Used in | Tokens | Auto-check | Licence | Updated |
|---|---|---|---|---|---|---|---|---|
| 7 | Replay-first debug flow for SGLang serving problems. An agent skill from sgl-project/sglang. | sgl-project/ | 37k | 3 repos | ~2.1k | Automated safety check: Pass | Apache-2.0 | today |
| 8 | Unified LLM torch-profiler triage skill for sglang, vllm, TensorRT-LLM, and TokenSpeed. | sgl-project/ | 37k | 2 repos | ~6.4k | Automated safety check: Pass | Apache-2.0 | today |
| 9 | Start and persistently pursue a goal to babysit an SGLang pull request until selected GitHub Actions workflows pass on the latest PR head. | sgl-project/ | 37k | 2 repos | ~3k | Automated safety check: Pass | Apache-2.0 | today |
| 10 | Debug hanging issues in SGLang distributed inference (TP/PP/DP/EP). | sgl-project/ | 37k | 2 repos | ~2.4k | Automated safety check: Pass | Apache-2.0 | today |
| 11 | Conventions for SGLang environment variables — where to define, how to access, how to name, and how to deprecate. | sgl-project/ | 37k | 2 repos | ~2.9k | Automated safety check: Pass | Apache-2.0 | today |
| 12 | Write, calibrate, and debug the prefill-vs-decode logprob (KL) consistency tests in sglang -- the two independent conditions a zero requires (every operator batch-invariant, and the two paths… | sgl-project/ | 37k | 2 repos | ~3.7k | Automated safety check: Pass | Apache-2.0 | today |
| 13 | Plans and audits Day-0 SGLang support for a new model release: scope, architecture gaps, PR order, validation gates and sanitized public evidence. | BBuf/ | 900 | — | ~2.3k | Automated safety check: Pass | No licence | 2 days ago |
| 14 | Code style for SGLang large classes Scheduler, TokenizerManager, and ModelRunner: frozen-code conventions and init orchestration style. | sgl-project/ | 37k | 2 repos | ~1.9k | Automated safety check: Pass | Apache-2.0 | today |
| 15 | Use with the dstack skill for model-serving work when the image, serving command, resources, backend/fleet choice, or service behavior is not proven. | dstackai/ | 2.3k | — | ~1.6k | Automated safety check: Pass | MPL-2.0 | today |
| 16 | Analyze, benchmark, diagnose, and optimize large-model inference deployments from hardware inventory, model details, workload traces, and latency or throughput SLOs. | rednote-machine-learning/ | 142 | — | ~4.5k | Automated safety check: Pass | Apache-2.0 | 1 mo ago |
| 17 | Reads SGLang or vLLM startup logs to show where GPU memory went and estimates how many concurrent requests fit at common token lengths. | BBuf/ | 900 | — | ~2.5k | Automated safety check: Pass | No licence | 2 days ago |
| 18 | 18.Graphsignal Profile AI inference workloads (vLLM, SGLang, TensorRT-LLM, PyTorch, any GPU application) with the Graphsignal profiler and read the results from its local /signals JSON endpoint. | graphsignal/ | 257 | — | ~6.2k | Automated safety check: Pass | Apache-2.0 | 9 days ago |
| 19 | Clean up noisy startup warnings and spurious prints in SGLang server logs. | guqiong96/ | 143 | 1 repo | ~4.5k | Automated safety check: Pass | Apache-2.0 | 2 days ago |
| 20 | Trigger the bot-cherry-pick workflow for a batch of merged PRs onto a release branch and monitor each run to completion. | sgl-project/ | 37k | 2 repos | ~4k | Automated safety check: Pass | Apache-2.0 | today |
| 21 | 21.One Eval 驱动 One-Eval 对 API 或本地模型做端到端评测,覆盖纯文本、多模态、代码生成、函数调用和 Agent benchmark。当用户想评测模型在一个或多个 benchmark 上的表现、比较分数、补充 metric,或生成图文评测报告时使用本 skill。 | OpenDCAI/ | 165 | — | ~2.4k | Automated safety check: Pass | Apache-2.0 | 1 mo ago |
| 22 | Analyzes Torch Profiler traces from SGLang, vLLM and TensorRT-LLM servers into kernel attribution, overlap and fusion tables. | BBuf/ | 900 | — | ~2.8k | Automated safety check: Pass | No licence | 2 days ago |
| 23 | Diagnose and correct GPT-QModel tokenizer initialization, tokenization normalization, special-token handling, prompt rendering, and chat-template problems. | ModelCloud/ | 1.3k | — | ~1.1k | Automated safety check: Pass | Unknown | today |
| 24 | A skill your agent uses when benchmarking denoise latency or profiling a diffusion bottleneck in SGLang. | sgl-project/ | 37k | 2 repos | ~2.4k | Automated safety check: Pass | Apache-2.0 | today |
| 25 | Looks up public original architecture diagrams for named LLM, vision-language, MoE, diffusion and OCR models and returns the image with its source attribution. | BBuf/ | 900 | — | ~1.2k | Automated safety check: Pass | No licence | 2 days ago |
| 26 | Fallback installer for milesdiffusion on a bare CUDA 12.9 Linux GPU box, reproducing the official radixark/milesdiffusion image's package versions and verifying them. | radixark/ | 107 | — | ~1.6k | Automated safety check: Pass | Apache-2.0 | today |
| 27 | A skill your agent uses when quantizing a diffusion DiT with NVIDIA ModelOpt and making the resulting FP8 or NVFP4 checkpoint loadable, verifiable, and benchmarkable in SGLang Diffusion. | sgl-project/ | 37k | 2 repos | ~5k | Automated safety check: Pass | Apache-2.0 | today |
| 28 | Guides reinforcement-learning post-training of LLMs with slime, which pairs Megatron-LM training with SGLang rollouts, including GRPO runs on GLM, Qwen3 and Llama 3 models. | Orchestra-Research/ | 13k | 5 repos | ~2.8k | Automated safety check: Pass | MIT | 3 mo ago |
| 29 | Step-by-step tutorial for adding a new lightweight JIT CUDA kernel to sglang's jitkernel module | guqiong96/ | 143 | 1 repo | ~10k | Automated safety check: Pass | Apache-2.0 | 2 days ago |
| 30 | Naming conventions for SGLang speculative decoding identifiers. | sgl-project/ | 37k | 2 repos | ~1.6k | Automated safety check: Pass | Apache-2.0 | today |
| 31 | Validate changed quantization math, packed formats, and inference kernels against a Torch oracle; adjudicate near-tie mismatches before rejecting optimized arithmetic. | ModelCloud/ | 1.3k | — | ~1.4k | Automated safety check: Pass | Unknown | today |
| 32 | Reviews SGLang changes the way its maintainers do, drawing on a bundled corpus of public PR review threads and a flowchart of how the diff runs. | BBuf/ | 900 | — | ~4.6k | Automated safety check: Pass | No licence | 2 days ago |
| 33 | Sets up a structured debugging session for a Dynamo bug — pull the report from a Linear ticket, GitHub issue, or pasted text, capture the environment, create a persistent worklog markdown file, and… | ai-dynamo/ | 8.2k | — | ~1.2k | Automated safety check: Pass | Apache-2.0 | today |
| 34 | Add a new model to the SGLang Cookbook (docs/, Mintlify), config-driven format — instantiate the model-agnostic template into a per-model config (+ benchmarks) JSX under src/snippets/configs/, an… | sgl-project/ | 37k | 2 repos | ~4k | Automated safety check: Pass | Apache-2.0 | today |
| 35 | Migrate a legacy-template SGLang cookbook page (monolithic per-model generator under docs/src/snippets/autoregressive/) onto the config-driven template (shared deployment.jsx / playground.jsx… | sgl-project/ | 37k | 2 repos | ~4k | Automated safety check: Pass | Apache-2.0 | today |
| 36 | Guide to SGLang CI workflow orchestration — stage ordering, fail-fast, gating, partitioning, execution modes, and debugging CI failures. | sgl-project/ | 37k | 2 repos | ~5.5k | Automated safety check: Pass | Apache-2.0 | today |
| 37 | A skill your agent uses when choosing the fastest SGLang Diffusion flags for a model, GPU, and VRAM budget. | sgl-project/ | 37k | 2 repos | ~13k | Automated safety check: Pass | Apache-2.0 | today |
| 38 | Guide for writing SGLang CI/UT tests. An agent skill from sgl-project/sglang. | sgl-project/ | 37k | 2 repos | ~5.2k | Automated safety check: Pass | Apache-2.0 | today |
| 39 | Covers serving LLMs with SGLang, whose RadixAttention reuses cached prefixes, and constraining output to JSON, regex or grammar for agent and tool-calling workloads. | Orchestra-Research/ | 13k | 3 repos | ~2.9k | Automated safety check: Pass | MIT | 3 mo ago |
| 40 | Trains LLMs with reinforcement learning using verl, from ByteDance's Seed team, with GRPO, PPO and other algorithms and swappable training and rollout backends. | Orchestra-Research/ | 13k | 3 repos | ~2.4k | Automated safety check: Pass | MIT | 3 mo ago |
| 41 | Step-by-step tutorial for adding a heavyweight AOT CUDA/C++ kernel to sgl-kernel (including tests & benchmarks) | sgl-project/ | 37k | 2 repos | ~3.4k | Automated safety check: Pass | Apache-2.0 | today |
| 42 | Review a pull request against the SGLang Cookbook (docs/, Mintlify) contribution checklist — the config-driven format (per-model config + benchmarks JSX consumed by the shared deployment.jsx /… | sgl-project/ | 37k | 2 repos | ~4.2k | Automated safety check: Pass | Apache-2.0 | today |
| 43 | Call this skill when you need to debug CUDA crashes in SGLang using kernel API logging | sgl-project/ | 37k | 2 repos | ~4.9k | Automated safety check: Pass | Apache-2.0 | today |
| 44 | Generate an e2e profiling trace of an SGLang server run. An agent skill from sgl-project/sglang. | sgl-project/ | 37k | 2 repos | ~1.1k | Automated safety check: Pass | Apache-2.0 | today |
| 45 | Requirements for the SGLang scripted runtime, chiefly when to add (vs not add) a harness API. | sgl-project/ | 37k | 2 repos | ~448 | Automated safety check: Pass | Apache-2.0 | today |
| 46 | Investigate consistently failing SGLang CI tests by extracting the failure signature from scheduled or rerun workflows, bisecting the passing/failing commit window, checking runner or hardware… | sgl-project/ | 37k | 2 repos | ~2.5k | Automated safety check: Pass | Apache-2.0 | today |
| 47 | A skill your agent uses when adding a new diffusion model or Diffusers pipeline to SGLang. | sgl-project/ | 37k | 2 repos | ~9.9k | Automated safety check: Pass | Apache-2.0 | today |
| 48 | 48.Recif Eval Run RecIF beam×concurrency eval and FlashRec vs SGLang/vLLM/TRT-LLM baselines. | sohu-mptc/ | 107 | — | ~915 | Automated safety check: Pass | Apache-2.0 | 5 days ago |
Questions, answered from the data.
What is the best SGLang skill?
SageMaker Serving Image Selection (official) from huggingface/skills ranks first of the 79 SGLang skills listed here, with the highest score: its repository has 11k GitHub stars, 1 other GitHub owner carry a copy, its SKILL.md loads about 4.6k tokens and it passes the automated safety check with no findings. Next come Debug Inference and Add Export Format.
Is there an official SGLang skill?
6 of the 79 SGLang skills are official, published by the vendor's own GitHub organization: SageMaker Serving Image Selection, Debug Inference, Add Export Format, Jetson Memory Audit, Jetson LLM Serve and 1 more.
How are these skills ranked?
By Skill Navigator score, which combines the GitHub stars of the skill's repository (shared across that repo's skills and discounted for large collections), how many other GitHub owners carry a copy of the skill, and automated SKILL.md quality checks, minus penalties for safety-check warnings and for each further skill from the same repository. Skills that fail the safety check are not listed.