Search
vLLM · Deployment
Skills
Sort:BestMost starsTrending todayTrending this weekTrending this monthNewestRecently updatedName
| # | Skill | Repository | Stars | Used in | Tokens | Auto-check | Licence | Updated |
|---|---|---|---|---|---|---|---|---|
| 1 | Chooses the right serving container and current image URI for deploying a Hugging Face model to a SageMaker endpoint, preferring Hugging Face images over generic ones. | huggingface/ | 11k | 1 repo | ~4.6k | Automated safety check: Pass | Apache-2.0 | 3 days ago |
| 2 | Deploys SageMaker endpoints with autoscaling, CloudWatch alarms and tags on by default, using scripts for real-time, scale-to-zero and async setups. | huggingface/ | 11k | 1 repo | ~6.9k | Automated safety check: Pass | Apache-2.0 | 3 days ago |
| 3 | Deploys LLMs with vLLM for high-throughput serving, covering the OpenAI-compatible server, offline batch inference, monitoring and a Docker rollout. | Orchestra-Research/ | 13k | 5 repos | ~2.3k | Automated safety check: Pass | MIT | 3 mo ago |
| 4 | Interactively build, push or load, and deploy an airunway component (controller or any provider) to the cluster | ai-runway/ | 102 | — | ~927 | Automated safety check: Pass | Apache-2.0 | 14 days ago |
| 5 | Deploy LLM models on OCI using AI Quick Actions (AQUA) - single model, multi-model, stacked (LoRA), with GPU shape selection, vLLM configuration, streaming, and tool calling. | oracle/ | 125 | — | ~2.4k | Automated safety check: Pass | UPL-1.0 | 1 mo ago |
| 6 | A skill your agent uses when adding, debugging, or validating a bring-your-own VLM in VSS RT-VLM, including custom Hugging Face or NGC checkpoints, vLLM adapters or plugins, model shims, and… | NVIDIA-AI-Blueprints/ | 1.9k | — | ~1.4k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 7 | Deploy vLLM to Kubernetes (K8s) with GPU support, health probes, and OpenAI-compatible API endpoint. | vllm-project/ | 102 | — | ~2k | Automated safety check: Pass | Apache-2.0 | 6 mo ago |
| 8 | Set up Prometheus and Grafana monitoring for AQUA vLLM model deployments on OCI. | oracle/ | 125 | — | ~1.5k | Automated safety check: Pass | UPL-1.0 | 1 mo ago |
| 9 | Implement a production SageMaker endpoint with autoscaling, CloudWatch alarms, and tags. | waybarrios/ | 534 | — | ~4.6k | Automated safety check: Pass | Apache-2.0 | 5 days ago |
| 10 | Diagnose and fix OCI AI Quick Actions (AQUA) issues including deployment failures, OOM errors, authorization problems, capacity issues, container errors, and policy misconfigurations. | oracle/ | 125 | — | ~1.8k | Automated safety check: Pass | UPL-1.0 | 1 mo ago |
| 11 | Select and verify the current region-specific serving container URI for a SageMaker model deployment. | waybarrios/ | 534 | — | ~4.3k | Automated safety check: Pass | Apache-2.0 | 5 days ago |
| 12 | Phase 1 of LLM deployment — for every leaf kernel × shape the model needs, verify numerical correctness on real NPU2 against the registry's GPU/vLLM-aligned standard. | Xilinx/ | 150 | — | ~3.4k | Automated safety check: Pass | MIT | yesterday |
| 13 | Generates and updates secure, production-ready Kubernetes YAML manifests optimized for GKE Autopilot and GKE Standard clusters. | google/ | 21k | — | ~3.1k | Automated safety check: Pass | Apache-2.0 | 2 days ago |
| 14 | How to swap the VLM in the VSS Alerts Blueprint — covers RTVI-VLM microservice deployment methods, all three VLM consumers (rtvi-vlm, vlm-as-verifier, vss-agent), and health checks. | NVIDIA/ | 3.6k | — | ~5k | Automated safety check: Notes | Apache-2.0 | 2 days ago |