Search
SGLang
Skills
Sort:BestMost starsTrending todayTrending this weekTrending this monthNewestRecently updatedName
| # | Skill | Repository | Stars | Used in | Tokens | Auto-check | Licence | Updated |
|---|---|---|---|---|---|---|---|---|
| 1 | Replay-first debug flow for SGLang serving problems. An agent skill from sgl-project/sglang. | sgl-project/ | 37k | 3 repos | ~2.1k | Automated safety check: Pass | Apache-2.0 | today |
| 2 | Unified LLM torch-profiler triage skill for sglang, vllm, TensorRT-LLM, and TokenSpeed. | sgl-project/ | 37k | 2 repos | ~6.4k | Automated safety check: Pass | Apache-2.0 | today |
| 3 | Start and persistently pursue a goal to babysit an SGLang pull request until selected GitHub Actions workflows pass on the latest PR head. | sgl-project/ | 37k | 2 repos | ~3k | Automated safety check: Pass | Apache-2.0 | today |
| 4 | Chooses the right serving container and current image URI for deploying a Hugging Face model to a SageMaker endpoint, preferring Hugging Face images over generic ones. | huggingface/ | 11k | 1 repo | ~4.6k | Automated safety check: Pass | Apache-2.0 | 2 days ago |
| 5 | Debug hanging issues in SGLang distributed inference (TP/PP/DP/EP). | sgl-project/ | 37k | 2 repos | ~2.4k | Automated safety check: Pass | Apache-2.0 | today |
| 6 | Conventions for SGLang environment variables — where to define, how to access, how to name, and how to deprecate. | sgl-project/ | 37k | 2 repos | ~2.9k | Automated safety check: Pass | Apache-2.0 | today |
| 7 | Write, calibrate, and debug the prefill-vs-decode logprob (KL) consistency tests in sglang -- the two independent conditions a zero requires (every operator batch-invariant, and the two paths… | sgl-project/ | 37k | 2 repos | ~3.7k | Automated safety check: Pass | Apache-2.0 | today |
| 8 | Plans and audits Day-0 SGLang support for a new model release: scope, architecture gaps, PR order, validation gates and sanitized public evidence. | BBuf/ | 938 | — | ~2.3k | Automated safety check: Pass | No licence | 5 days ago |
| 9 | Code style for SGLang large classes Scheduler, TokenizerManager, and ModelRunner: frozen-code conventions and init orchestration style. | sgl-project/ | 37k | 2 repos | ~1.9k | Automated safety check: Pass | Apache-2.0 | today |
| 10 | Use with the dstack skill for model-serving work when the image, serving command, resources, backend/fleet choice, or service behavior is not proven. | dstackai/ | 2.3k | — | ~1.6k | Automated safety check: Pass | MPL-2.0 | yesterday |
| 11 | Analyze, benchmark, diagnose, and optimize large-model inference deployments from hardware inventory, model details, workload traces, and latency or throughput SLOs. | rednote-machine-learning/ | 144 | — | ~4.5k | Automated safety check: Pass | Apache-2.0 | 1 mo ago |
| 12 | Debug inference clients that use an attached provider and its native endpoint, including hosted APIs and host-local Ollama, vLLM, SGLang, TRT-LLM, LM Studio, or NIM. | NVIDIA/ | 16k | — | ~1.9k | Automated safety check: Pass | Apache-2.0 | today |
| 13 | Reads SGLang or vLLM startup logs to show where GPU memory went and estimates how many concurrent requests fit at common token lengths. | BBuf/ | 938 | — | ~2.5k | Automated safety check: Pass | No licence | 5 days ago |
| 14 | 14.Graphsignal Profile AI inference workloads (vLLM, SGLang, TensorRT-LLM, PyTorch, any GPU application) with the Graphsignal profiler and read the results from its local /signals JSON endpoint. | graphsignal/ | 257 | — | ~6.3k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 15 | Clean up noisy startup warnings and spurious prints in SGLang server logs. | guqiong96/ | 144 | 1 repo | ~4.5k | Automated safety check: Pass | Apache-2.0 | 5 days ago |
| 16 | 16.One Eval 驱动 One-Eval 对 API 或本地模型做端到端评测,覆盖纯文本、多模态、代码生成、函数调用和 Agent benchmark。当用户想评测模型在一个或多个 benchmark 上的表现、比较分数、补充 metric,或生成图文评测报告时使用本 skill。 | OpenDCAI/ | 165 | — | ~2.4k | Automated safety check: Pass | Apache-2.0 | 1 mo ago |
| 17 | Trigger the bot-cherry-pick workflow for a batch of merged PRs onto a release branch and monitor each run to completion. | sgl-project/ | 37k | 2 repos | ~4k | Automated safety check: Pass | Apache-2.0 | today |
| 18 | Analyzes Torch Profiler traces from SGLang, vLLM and TensorRT-LLM servers into kernel attribution, overlap and fusion tables. | BBuf/ | 938 | — | ~2.8k | Automated safety check: Pass | No licence | 5 days ago |
| 19 | Diagnose and correct GPT-QModel tokenizer initialization, tokenization normalization, special-token handling, prompt rendering, and chat-template problems. | ModelCloud/ | 1.3k | — | ~1.1k | Automated safety check: Pass | Unknown | today |
| 20 | Add a new model export format to AutoRound (e.g., autoround, autogptq, autoawq, gguf, llmcompressor). | intel/ | 1.6k | — | ~1.9k | Automated safety check: Pass | Apache-2.0 | today |
| 21 | A skill your agent uses when benchmarking denoise latency or profiling a diffusion bottleneck in SGLang. | sgl-project/ | 37k | 2 repos | ~2.4k | Automated safety check: Pass | Apache-2.0 | today |
| 22 | Looks up public original architecture diagrams for named LLM, vision-language, MoE, diffusion and OCR models and returns the image with its source attribution. | BBuf/ | 938 | — | ~1.2k | Automated safety check: Pass | No licence | 5 days ago |
| 23 | Fallback installer for milesdiffusion on a bare CUDA 12.9 Linux GPU box, reproducing the official radixark/milesdiffusion image's package versions and verifying them. | radixark/ | 109 | — | ~1.6k | Automated safety check: Pass | Apache-2.0 | today |
| 24 | A skill your agent uses when quantizing a diffusion DiT with NVIDIA ModelOpt and making the resulting FP8 or NVFP4 checkpoint loadable, verifiable, and benchmarkable in SGLang Diffusion. | sgl-project/ | 37k | 2 repos | ~5k | Automated safety check: Pass | Apache-2.0 | today |
| 25 | Step-by-step tutorial for adding a new lightweight JIT CUDA kernel to sglang's jitkernel module | guqiong96/ | 144 | 1 repo | ~10k | Automated safety check: Pass | Apache-2.0 | 5 days ago |
| 26 | Naming conventions for SGLang speculative decoding identifiers. | sgl-project/ | 37k | 2 repos | ~1.6k | Automated safety check: Pass | Apache-2.0 | today |
| 27 | Validate changed quantization math, packed formats, and inference kernels against a Torch oracle; adjudicate near-tie mismatches before rejecting optimized arithmetic. | ModelCloud/ | 1.3k | — | ~1.4k | Automated safety check: Pass | Unknown | today |
| 28 | Guides reinforcement-learning post-training of LLMs with slime, which pairs Megatron-LM training with SGLang rollouts, including GRPO runs on GLM, Qwen3 and Llama 3 models. | Orchestra-Research/ | 13k | 4 repos | ~2.8k | Automated safety check: Pass | MIT | 3 mo ago |
| 29 | Reviews SGLang changes the way its maintainers do, drawing on a bundled corpus of public PR review threads and a flowchart of how the diff runs. | BBuf/ | 938 | — | ~4.6k | Automated safety check: Pass | No licence | 5 days ago |
| 30 | Sets up a structured debugging session for a Dynamo bug — pull the report from a Linear ticket, GitHub issue, or pasted text, capture the environment, create a persistent worklog markdown file, and… | ai-dynamo/ | 8.3k | — | ~1.2k | Automated safety check: Pass | Apache-2.0 | today |
| 31 | Add a new model to the SGLang Cookbook (docs/, Mintlify), config-driven format — instantiate the model-agnostic template into a per-model config (+ benchmarks) JSX under src/snippets/configs/, an… | sgl-project/ | 37k | 2 repos | ~4k | Automated safety check: Pass | Apache-2.0 | today |
| 32 | Migrate a legacy-template SGLang cookbook page (monolithic per-model generator under docs/src/snippets/autoregressive/) onto the config-driven template (shared deployment.jsx / playground.jsx… | sgl-project/ | 37k | 2 repos | ~4k | Automated safety check: Pass | Apache-2.0 | today |
| 33 | Guide to SGLang CI workflow orchestration — stage ordering, fail-fast, gating, partitioning, execution modes, and debugging CI failures. | sgl-project/ | 37k | 2 repos | ~5.5k | Automated safety check: Pass | Apache-2.0 | today |
| 34 | A skill your agent uses when choosing the fastest SGLang Diffusion flags for a model, GPU, and VRAM budget. | sgl-project/ | 37k | 2 repos | ~13k | Automated safety check: Pass | Apache-2.0 | today |
| 35 | Guide for writing SGLang CI/UT tests. An agent skill from sgl-project/sglang. | sgl-project/ | 37k | 2 repos | ~5.2k | Automated safety check: Pass | Apache-2.0 | today |
| 36 | Step-by-step tutorial for adding a heavyweight AOT CUDA/C++ kernel to sgl-kernel (including tests & benchmarks) | sgl-project/ | 37k | 2 repos | ~3.4k | Automated safety check: Pass | Apache-2.0 | today |
| 37 | Review a pull request against the SGLang Cookbook (docs/, Mintlify) contribution checklist — the config-driven format (per-model config + benchmarks JSX consumed by the shared deployment.jsx /… | sgl-project/ | 37k | 2 repos | ~4.2k | Automated safety check: Pass | Apache-2.0 | today |
| 38 | Call this skill when you need to debug CUDA crashes in SGLang using kernel API logging | sgl-project/ | 37k | 2 repos | ~4.9k | Automated safety check: Pass | Apache-2.0 | today |
| 39 | Generate an e2e profiling trace of an SGLang server run. An agent skill from sgl-project/sglang. | sgl-project/ | 37k | 2 repos | ~1.1k | Automated safety check: Pass | Apache-2.0 | today |
| 40 | Requirements for the SGLang scripted runtime, chiefly when to add (vs not add) a harness API. | sgl-project/ | 37k | 2 repos | ~448 | Automated safety check: Pass | Apache-2.0 | today |
| 41 | Investigate consistently failing SGLang CI tests by extracting the failure signature from scheduled or rerun workflows, bisecting the passing/failing commit window, checking runner or hardware… | sgl-project/ | 37k | 2 repos | ~2.5k | Automated safety check: Pass | Apache-2.0 | today |
| 42 | A skill your agent uses when adding a new diffusion model or Diffusers pipeline to SGLang. | sgl-project/ | 37k | 2 repos | ~9.9k | Automated safety check: Pass | Apache-2.0 | today |
| 43 | 43.Recif Eval Run RecIF beam×concurrency eval and FlashRec vs SGLang/vLLM/TRT-LLM baselines. | sohu-mptc/ | 107 | — | ~915 | Automated safety check: Pass | Apache-2.0 | today |
| 44 | 44.Dev Bump Bump Heron version via the VERSION-file SSOT. An agent skill from Netis/heron. | Netis/ | 102 | — | ~983 | Automated safety check: Pass | Apache-2.0 | 6 days ago |
| 45 | Covers serving LLMs with SGLang, whose RadixAttention reuses cached prefixes, and constraining output to JSON, regex or grammar for agent and tool-calling workloads. | Orchestra-Research/ | 13k | 2 repos | ~2.9k | Automated safety check: Pass | MIT | 3 mo ago |
| 46 | Trains LLMs with reinforcement learning using verl, from ByteDance's Seed team, with GRPO, PPO and other algorithms and swappable training and rollout backends. | Orchestra-Research/ | 13k | 2 repos | ~2.4k | Automated safety check: Pass | MIT | 3 mo ago |
| 47 | How SGLang's runtime configuration and process-global state are organized (RuntimeContext tiers, publish + namespace config bags, the pristine ServerArgs seed, override entry points… | sgl-project/ | 37k | 2 repos | ~11k | Automated safety check: Pass | Apache-2.0 | today |
| 48 | Run and troubleshoot the inference backends on the DGX Spark pair that FastLLM proxies to — starting or stopping models with vLLM, SGLang or sparkrun, choosing memory and speculative-decoding… | azrtydxb/ | 108 | — | ~980 | Automated safety check: Pass | Apache-2.0 | 5 days ago |