Search
vLLM · For data scientists
Skills
Sort:BestMost starsTrending todayTrending this weekTrending this monthNewestRecently updatedName
| # | Skill | Repository | Stars | Used in | Tokens | Auto-check | Licence | Updated |
|---|---|---|---|---|---|---|---|---|
| 1 | Runs lm-evaluation-harness to benchmark language models on academic suites such as MMLU, GSM8K and HumanEval, compare models and track training checkpoints. | Orchestra-Research/ | 13k | 8 repos | ~3k | Automated safety check: Pass | MIT | 3 mo ago |
| 2 | Chooses the right serving container and current image URI for deploying a Hugging Face model to a SageMaker endpoint, preferring Hugging Face images over generic ones. | huggingface/ | 11k | 1 repo | ~4.6k | Automated safety check: Pass | Apache-2.0 | 2 days ago |
| 3 | Runs evaluations of Hugging Face Hub models on local hardware with inspect-ai or lighteval, and helps choose between vLLM, Transformers and accelerate backends. | huggingface/ | 11k | 2 repos | ~1.6k | Automated safety check: Pass | Apache-2.0 | 2 days ago |
| 4 | Fetch and diagnose vLLM Buildkite CI failure logs. An agent skill from guqiong96/Lvllm. | guqiong96/ | 465 | 2 repos | ~349 | Automated safety check: Pass | Apache-2.0 | 19 days ago |
| 5 | Low-token Codex session/thread title organizer. An agent skill from David-Lzy/codex_session_renamer. | David-Lzy/ | 97 | — | ~1.2k | Automated safety check: Pass | MIT | 3 mo ago |
| 6 | Diagnose and correct GPT-QModel tokenizer initialization, tokenization normalization, special-token handling, prompt rendering, and chat-template problems. | ModelCloud/ | 1.3k | — | ~1.1k | Automated safety check: Pass | Unknown | today |
| 7 | Add a new diffusion model (text-to-image, text-to-video, image-to-video, text-to-audio, image editing) to vLLM-Omni, including native non-Diffusers ports, reference-parity validation, Cache-DiT… | vllm-project/ | 7.1k | — | ~7k | Automated safety check: Pass | Apache-2.0 | today |
| 8 | Plans memory headroom, works through out-of-memory failures and watches temperature and power during long ML training jobs on NVIDIA DGX Spark. | wshobson/ | 40k | — | ~2k | Automated safety check: Pass | MIT | 6 days ago |
| 9 | Review and upgrade MetaX model support against a target vLLM revision and installed MACA components, recursively including model-dependent attention and kernels. | MetaX-MACA/ | 180 | — | ~3.2k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 10 | 10.Add Recipe Add or update an in-repository vLLM-Omni model recipe with verified task, input, output, hardware, command, feature, and validation contracts. | vllm-project/ | 7.1k | — | ~1.4k | Automated safety check: Pass | Apache-2.0 | today |
| 11 | Quick-reference help for fine-tuning language models with Axolotl, covering YAML configs, FSDP, context parallelism, compressed saves and dataset formats. | Orchestra-Research/ | 13k | 8 repos | ~1.2k | Automated safety check: Pass | MIT | 3 mo ago |
| 12 | Write or review Triton kernels for vLLM, with practical guidance for generated-code inspection, launch grids, indexing, specialization, tuning, and representative performance validation. | guqiong96/ | 465 | 1 repo | ~831 | Automated safety check: Pass | Apache-2.0 | 19 days ago |
| 13 | Review and adapt vllmmetax/registry registrations, quantization configurations, CustomOps and kernel dispatch against a target vLLM revision and installed MetaX APIs. | MetaX-MACA/ | 180 | — | ~1.8k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 14 | Query Finelog logs and telemetry for Iris tasks, workers, profiles, training, vLLM, and cross-cluster forwarding. | marin-community/ | 3.9k | — | ~1.1k | Automated safety check: Pass | Apache-2.0 | today |
| 15 | Produces ranked, evidence-grounded next steps when an ML experiment has stalled, drawing on a Leeroopedia knowledge base or on fetched docs and issues. | Leeroo-AI/ | 195 | — | ~4.8k | Automated safety check: Pass | Apache-2.0 | 6 mo ago |
| 16 | Review PRs and local changes for hsliuustc0106/vllm-rlt: Ouro engine and KV correctness, serving behavior, and BF16 accuracy/speed evidence. | ThinkFlowLab/ | 149 | — | ~1.1k | Automated safety check: Pass | Apache-2.0 | today |
| 17 | Find evidence-backed simplification candidates in vLLM-Omni and, when requested, turn them into focused proposals or code changes. | vllm-project/ | 7.1k | — | ~2.7k | Automated safety check: Pass | Apache-2.0 | today |
| 18 | Checks training code, configs and math against documented framework behavior before an expensive run, citing a knowledge base or official docs for every claim. | Leeroo-AI/ | 195 | — | ~3.8k | Automated safety check: Pass | Apache-2.0 | 6 mo ago |
| 19 | Run, resume, monitor, diagnose, and report Quark Quant-Perf workflows for PyTorch and HuggingFace transformers models. | amd/ | 182 | — | ~3k | Automated safety check: Pass | MIT | 13 days ago |
| 20 | Activation-aware weight quantization for 4-bit LLM compression with 3x speedup and minimal accuracy loss. | Orchestra-Research/ | 13k | 2 repos | ~2.1k | Automated safety check: Pass | MIT | 3 mo ago |
| 21 | High-performance RLHF framework with Ray+vLLM acceleration. An agent skill from Orchestra-Research/AI-Research-SKILLs. | Orchestra-Research/ | 13k | 2 repos | ~2.1k | Automated safety check: Notes | MIT | 3 mo ago |
| 22 | Covers serving LLMs with SGLang, whose RadixAttention reuses cached prefixes, and constraining output to JSON, regex or grammar for agent and tool-calling workloads. | Orchestra-Research/ | 13k | 2 repos | ~2.9k | Automated safety check: Pass | MIT | 3 mo ago |
| 23 | Guides reinforcement-learning research with torchforge, Meta's PyTorch-native library that keeps RL algorithms apart from infrastructure, including GRPO math-reasoning runs. | Orchestra-Research/ | 13k | 2 repos | ~2.5k | Automated safety check: Pass | MIT | 3 mo ago |
| 24 | Trains LLMs with reinforcement learning using verl, from ByteDance's Seed team, with GRPO, PPO and other algorithms and swappable training and rollout backends. | Orchestra-Research/ | 13k | 2 repos | ~2.4k | Automated safety check: Pass | MIT | 3 mo ago |
| 25 | Review and adapt MetaX attention backends, MLA, sparse indexers, cache layouts, and their kernel wrappers against a target vLLM revision and the actually installed MetaX component APIs. | MetaX-MACA/ | 180 | — | ~2.3k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 26 | 26.Rlt Perf Opt Analyze and optimize vllm-rlt inference performance using reproducible unprofiled benchmarks, paired ops-only/full profiles, source-level attribution, and correctness checks. | ThinkFlowLab/ | 149 | — | ~4.6k | Automated safety check: Pass | Apache-2.0 | today |
| 27 | Audit and adapt monkey patches in vllmmetax/patch/ against a target upstream revision. | MetaX-MACA/ | 180 | — | ~2.8k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 28 | Establish the shared environment, source/runtime correspondence, target confirmation and validation evidence for MetaX vLLM upgrades. | MetaX-MACA/ | 180 | — | ~1.1k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 29 | This is a skill for benchmarking the efficiency of automatic prefix caching in vLLM using fixed prompts, real-world datasets, or synthetic prefix/suffix patterns. | vllm-project/ | 102 | — | ~1.4k | Automated safety check: Pass | Apache-2.0 | 6 mo ago |
| 30 | DESIGN a non-trivial codebase change (Harbor / MarinSkyRL / vLLM / OT-Agent / LLaMA-Factory) as a dependency-ordered STAGED PLAN before writing code — a feature port, a multi-step fix with parity… | open-thoughts/ | 301 | — | ~1.5k | Automated safety check: Pass | Apache-2.0 | 12 days ago |
| 31 | Trim large MetaX model directories for dummy smoke tests or real-checkpoint loading on limited GPUs. | MetaX-MACA/ | 180 | — | ~2.4k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 32 | Diagnose and fix OCI AI Quick Actions (AQUA) issues including deployment failures, OOM errors, authorization problems, capacity issues, container errors, and policy misconfigurations. | oracle/ | 125 | — | ~1.8k | Automated safety check: Pass | UPL-1.0 | 1 mo ago |
| 33 | Upgrade vllm-omni NPU model runners (OmniNPUModelRunner, NPUARModelRunner, NPUGenerationModelRunner) to align with the latest vllm-ascend NPUModelRunner while preserving omni-specific logic. | vllm-project/ | 7.1k | — | ~3k | Automated safety check: Pass | Apache-2.0 | today |
| 34 | Given a .mlir file (or a directory of .mlir files) with TTIR ops, run the same TTIR normalization passes as D2MFrontendPipeline before D2M, then produce per-file outputs: preprocessed.mlir… | tenstorrent/ | 314 | — | ~1.8k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 35 | Profile vLLM and RT-VLM inference to identify and remove GPU-idle gaps, underfilled batches, transfer stalls, serialized multimodal work, scheduler gaps, or KV pressure. | NVIDIA-AI-Blueprints/ | 1.9k | — | ~1.5k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 36 | Phase 1 of LLM deployment — for every leaf kernel × shape the model needs, verify numerical correctness on real NPU2 against the registry's GPU/vLLM-aligned standard. | Xilinx/ | 150 | — | ~3.4k | Automated safety check: Pass | MIT | yesterday |
| 37 | EXECUTE a staged codebase plan (from code-create-staged-plan or an existing notes/<codebase/ plan) one stage at a time, gate-by-gate, while keeping the local clone ground truth and a dated… | open-thoughts/ | 301 | — | ~977 | Automated safety check: Pass | Apache-2.0 | 12 days ago |
| 38 | Interactive online benchmark orchestrator for vLLM inference services using vllm bench serve. | ascend-ai-coding/ | 174 | — | ~5.6k | Automated safety check: Pass | No licence | yesterday |
| 39 | Machine-learning research loop for dataset curation, fine-tuning, evaluation, inference deployment, experiment tracking, and model explainability. | AnastasiyaW/ | 154 | — | ~794 | Automated safety check: Pass | MIT | yesterday |
| 40 | Build and operate Retrieval-Augmented Generation (RAG) infrastructure with vector stores, embedding pipelines, and hybrid search. | BagelHole/ | 1.2k | — | ~2.1k | Automated safety check: Pass | MIT | 4 mo ago |
| 41 | 41.Ais Bench AISBench Benchmark - AI model evaluation tool for Ascend NPU. | ascend-ai-coding/ | 174 | — | ~2.7k | Automated safety check: Pass | No licence | yesterday |
| 42 | 42.Open Weights A skill your agent uses when choosing an open-weight LLM and clearing it for use — which family and size fit the task, the hardware and the budget, and above all whether the license permits shipping. | ericrisco/ | 180 | — | ~4.1k | Automated safety check: Pass | MIT | yesterday |
| 43 | Plan and execute production ML engineering work — model training and fine-tuning (LoRA/QLoRA), evaluation and eval-set design, quantization decisions, inference deployment, lineage, feature parity… | magnus919/ | 115 | — | ~1.5k | Automated safety check: Pass | MIT | yesterday |