Search

SGLang

80 skills found.
Search results
#SkillRepositoryStarsUsed inTokensAuto-checkLicenceUpdated
1

Replay-first debug flow for SGLang serving problems. An agent skill from sgl-project/sglang.

sgl-project/sglang37k3 repos~2.1kAutomated safety check: PassApache-2.0today
2

Unified LLM torch-profiler triage skill for sglang, vllm, TensorRT-LLM, and TokenSpeed.

sgl-project/sglang37k2 repos~6.4kAutomated safety check: PassApache-2.0today
3

Start and persistently pursue a goal to babysit an SGLang pull request until selected GitHub Actions workflows pass on the latest PR head.

sgl-project/sglang37k2 repos~3kAutomated safety check: PassApache-2.0today
4

Chooses the right serving container and current image URI for deploying a Hugging Face model to a SageMaker endpoint, preferring Hugging Face images over generic ones.

huggingface/skills11k1 repo~4.6kAutomated safety check: PassApache-2.02 days ago
5

Debug hanging issues in SGLang distributed inference (TP/PP/DP/EP).

sgl-project/sglang37k2 repos~2.4kAutomated safety check: PassApache-2.0today
6

Conventions for SGLang environment variables — where to define, how to access, how to name, and how to deprecate.

sgl-project/sglang37k2 repos~2.9kAutomated safety check: PassApache-2.0today
7

Write, calibrate, and debug the prefill-vs-decode logprob (KL) consistency tests in sglang -- the two independent conditions a zero requires (every operator batch-invariant, and the two paths…

sgl-project/sglang37k2 repos~3.7kAutomated safety check: PassApache-2.0today
8

Plans and audits Day-0 SGLang support for a new model release: scope, architecture gaps, PR order, validation gates and sanitized public evidence.

BBuf/AI-Infra-Auto-Driven-SKILLS938—~2.3kAutomated safety check: PassNo licence5 days ago
9

Code style for SGLang large classes Scheduler, TokenizerManager, and ModelRunner: frozen-code conventions and init orchestration style.

sgl-project/sglang37k2 repos~1.9kAutomated safety check: PassApache-2.0today
10

Use with the dstack skill for model-serving work when the image, serving command, resources, backend/fleet choice, or service behavior is not proven.

dstackai/dstack2.3k—~1.6kAutomated safety check: PassMPL-2.0yesterday
11

Analyze, benchmark, diagnose, and optimize large-model inference deployments from hardware inventory, model details, workload traces, and latency or throughput SLOs.

rednote-machine-learning/Inference-autopilot144—~4.5kAutomated safety check: PassApache-2.01 mo ago
12
12.Debug InferenceOfficial

Debug inference clients that use an attached provider and its native endpoint, including hosted APIs and host-local Ollama, vLLM, SGLang, TRT-LLM, LM Studio, or NIM.

NVIDIA/OpenShell16k—~1.9kAutomated safety check: PassApache-2.0today
13

Reads SGLang or vLLM startup logs to show where GPU memory went and estimates how many concurrent requests fit at common token lengths.

BBuf/AI-Infra-Auto-Driven-SKILLS938—~2.5kAutomated safety check: PassNo licence5 days ago
14

Profile AI inference workloads (vLLM, SGLang, TensorRT-LLM, PyTorch, any GPU application) with the Graphsignal profiler and read the results from its local /signals JSON endpoint.

graphsignal/graphsignal257—~6.3kAutomated safety check: PassApache-2.0yesterday
15

Clean up noisy startup warnings and spurious prints in SGLang server logs.

guqiong96/Lsglang1441 repo~4.5kAutomated safety check: PassApache-2.05 days ago
16

驱动 One-Eval 对 API 或本地模型做端到端评测,覆盖纯文本、多模态、代码生成、函数调用和 Agent benchmark。当用户想评测模型在一个或多个 benchmark 上的表现、比较分数、补充 metric,或生成图文评测报告时使用本 skill。

OpenDCAI/One-Eval165—~2.4kAutomated safety check: PassApache-2.01 mo ago
17

Trigger the bot-cherry-pick workflow for a batch of merged PRs onto a release branch and monitor each run to completion.

sgl-project/sglang37k2 repos~4kAutomated safety check: PassApache-2.0today
18

Analyzes Torch Profiler traces from SGLang, vLLM and TensorRT-LLM servers into kernel attribution, overlap and fusion tables.

BBuf/AI-Infra-Auto-Driven-SKILLS938—~2.8kAutomated safety check: PassNo licence5 days ago
19

Diagnose and correct GPT-QModel tokenizer initialization, tokenization normalization, special-token handling, prompt rendering, and chat-template problems.

ModelCloud/GPTQModel1.3k—~1.1kAutomated safety check: PassUnknowntoday
20

Add a new model export format to AutoRound (e.g., autoround, autogptq, autoawq, gguf, llmcompressor).

intel/auto-round1.6k—~1.9kAutomated safety check: PassApache-2.0today
21

A skill your agent uses when benchmarking denoise latency or profiling a diffusion bottleneck in SGLang.

sgl-project/sglang37k2 repos~2.4kAutomated safety check: PassApache-2.0today
22

Looks up public original architecture diagrams for named LLM, vision-language, MoE, diffusion and OCR models and returns the image with its source attribution.

BBuf/AI-Infra-Auto-Driven-SKILLS938—~1.2kAutomated safety check: PassNo licence5 days ago
23

Fallback installer for milesdiffusion on a bare CUDA 12.9 Linux GPU box, reproducing the official radixark/milesdiffusion image's package versions and verifying them.

radixark/miles_diffusion109—~1.6kAutomated safety check: PassApache-2.0today
24

A skill your agent uses when quantizing a diffusion DiT with NVIDIA ModelOpt and making the resulting FP8 or NVFP4 checkpoint loadable, verifiable, and benchmarkable in SGLang Diffusion.

sgl-project/sglang37k2 repos~5kAutomated safety check: PassApache-2.0today
25

Step-by-step tutorial for adding a new lightweight JIT CUDA kernel to sglang's jitkernel module

guqiong96/Lsglang1441 repo~10kAutomated safety check: PassApache-2.05 days ago
26

Naming conventions for SGLang speculative decoding identifiers.

sgl-project/sglang37k2 repos~1.6kAutomated safety check: PassApache-2.0today
27

Validate changed quantization math, packed formats, and inference kernels against a Torch oracle; adjudicate near-tie mismatches before rejecting optimized arithmetic.

ModelCloud/GPTQModel1.3k—~1.4kAutomated safety check: PassUnknowntoday
28

Guides reinforcement-learning post-training of LLMs with slime, which pairs Megatron-LM training with SGLang rollouts, including GRPO runs on GLM, Qwen3 and Llama 3 models.

Orchestra-Research/AI-Research-SKILLs13k4 repos~2.8kAutomated safety check: PassMIT3 mo ago
29

Reviews SGLang changes the way its maintainers do, drawing on a bundled corpus of public PR review threads and a flowchart of how the diff runs.

BBuf/AI-Infra-Auto-Driven-SKILLS938—~4.6kAutomated safety check: PassNo licence5 days ago
30

Sets up a structured debugging session for a Dynamo bug — pull the report from a Linear ticket, GitHub issue, or pasted text, capture the environment, create a persistent worklog markdown file, and…

ai-dynamo/dynamo8.3k—~1.2kAutomated safety check: PassApache-2.0today
31

Add a new model to the SGLang Cookbook (docs/, Mintlify), config-driven format — instantiate the model-agnostic template into a per-model config (+ benchmarks) JSX under src/snippets/configs/, an…

sgl-project/sglang37k2 repos~4kAutomated safety check: PassApache-2.0today
32

Migrate a legacy-template SGLang cookbook page (monolithic per-model generator under docs/src/snippets/autoregressive/) onto the config-driven template (shared deployment.jsx / playground.jsx…

sgl-project/sglang37k2 repos~4kAutomated safety check: PassApache-2.0today
33

Guide to SGLang CI workflow orchestration — stage ordering, fail-fast, gating, partitioning, execution modes, and debugging CI failures.

sgl-project/sglang37k2 repos~5.5kAutomated safety check: PassApache-2.0today
34

A skill your agent uses when choosing the fastest SGLang Diffusion flags for a model, GPU, and VRAM budget.

sgl-project/sglang37k2 repos~13kAutomated safety check: PassApache-2.0today
35

Guide for writing SGLang CI/UT tests. An agent skill from sgl-project/sglang.

sgl-project/sglang37k2 repos~5.2kAutomated safety check: PassApache-2.0today
36

Step-by-step tutorial for adding a heavyweight AOT CUDA/C++ kernel to sgl-kernel (including tests & benchmarks)

sgl-project/sglang37k2 repos~3.4kAutomated safety check: PassApache-2.0today
37

Review a pull request against the SGLang Cookbook (docs/, Mintlify) contribution checklist — the config-driven format (per-model config + benchmarks JSX consumed by the shared deployment.jsx /…

sgl-project/sglang37k2 repos~4.2kAutomated safety check: PassApache-2.0today
38

Call this skill when you need to debug CUDA crashes in SGLang using kernel API logging

sgl-project/sglang37k2 repos~4.9kAutomated safety check: PassApache-2.0today
39

Generate an e2e profiling trace of an SGLang server run. An agent skill from sgl-project/sglang.

sgl-project/sglang37k2 repos~1.1kAutomated safety check: PassApache-2.0today
40

Requirements for the SGLang scripted runtime, chiefly when to add (vs not add) a harness API.

sgl-project/sglang37k2 repos~448Automated safety check: PassApache-2.0today
41

Investigate consistently failing SGLang CI tests by extracting the failure signature from scheduled or rerun workflows, bisecting the passing/failing commit window, checking runner or hardware…

sgl-project/sglang37k2 repos~2.5kAutomated safety check: PassApache-2.0today
42

A skill your agent uses when adding a new diffusion model or Diffusers pipeline to SGLang.

sgl-project/sglang37k2 repos~9.9kAutomated safety check: PassApache-2.0today
43

Run RecIF beam×concurrency eval and FlashRec vs SGLang/vLLM/TRT-LLM baselines.

sohu-mptc/FlashRec107—~915Automated safety check: PassApache-2.0today
44

Bump Heron version via the VERSION-file SSOT. An agent skill from Netis/heron.

Netis/heron102—~983Automated safety check: PassApache-2.06 days ago
45

Covers serving LLMs with SGLang, whose RadixAttention reuses cached prefixes, and constraining output to JSON, regex or grammar for agent and tool-calling workloads.

Orchestra-Research/AI-Research-SKILLs13k2 repos~2.9kAutomated safety check: PassMIT3 mo ago
46

Trains LLMs with reinforcement learning using verl, from ByteDance's Seed team, with GRPO, PPO and other algorithms and swappable training and rollout backends.

Orchestra-Research/AI-Research-SKILLs13k2 repos~2.4kAutomated safety check: PassMIT3 mo ago
47

How SGLang's runtime configuration and process-global state are organized (RuntimeContext tiers, publish + namespace config bags, the pristine ServerArgs seed, override entry points…

sgl-project/sglang37k2 repos~11kAutomated safety check: PassApache-2.0today
48

Run and troubleshoot the inference backends on the DGX Spark pair that FastLLM proxies to — starting or stopping models with vLLM, SGLang or sparkrun, choosing memory and speculative-decoding…

azrtydxb/Fastllm-proxy108—~980Automated safety check: PassApache-2.05 days ago