Search
NVIDIA AI Platform
Skills
Sort:BestMost starsTrending todayTrending this weekTrending this monthNewestRecently updatedName
| # | Skill | Repository | Stars | Used in | Tokens | Auto-check | Licence | Updated |
|---|---|---|---|---|---|---|---|---|
| 97 | Estimate GPU memory usage for Megatron-based MoE (Mixture of Experts) and dense models. | yzlnew/ | 149 | — | ~2.2k | Automated safety check: Pass | No licence | 3 mo ago |
| 98 | Evaluates LLMs across 100+ benchmarks from 18+ harnesses (MMLU, HumanEval, GSM8K, safety, VLM) with multi-backend execution. | Orchestra-Research/ | 13k | 2 repos | ~3.1k | Automated safety check: Pass | MIT | 3 mo ago |
| 99 | Covers serving LLMs with SGLang, whose RadixAttention reuses cached prefixes, and constraining output to JSON, regex or grammar for agent and tool-calling workloads. | Orchestra-Research/ | 13k | 2 repos | ~2.9k | Automated safety check: Pass | MIT | 3 mo ago |
| 100 | Sets up large-scale LLM training with NVIDIA Megatron-Core, choosing tensor, pipeline, data, context and expert parallelism for a given model size and GPU count. | Orchestra-Research/ | 13k | 2 repos | ~2.4k | Automated safety check: Pass | MIT | 3 mo ago |
| 101 | Detects CPU, GPU, memory and disk resources before heavy scientific tasks and writes a JSON file with advice on parallelism, out-of-core work and GPU use. | davila7/ | 33k | 10 repos | ~2.4k | Automated safety check: Pass | MIT | today |
| 102 | 102.Init GPU Server Initialize a Draw Things GPU server with GPUScript, including script sync, Docker/CUDA/NVIDIA runtime setup, 7T data disk mounting, mergerfs, and end-to-end GPU verification. | drawthingsai/ | 584 | — | ~2.2k | Automated safety check: Pass | GPL-3.0 | yesterday |
| 103 | 103.Nvwb Manage NVIDIA AI Workbench projects, contexts, builds, and environments via the nvwb CLI. | brevdev/ | 146 | — | ~1.1k | Automated safety check: Pass | Apache-2.0 | 3 days ago |
| 104 | Runs the Megatron-LM autoformat script and its linting tools before a pull request, and keeps Python imports in order with isort. | NVIDIA/ | 18k | — | ~316 | Automated safety check: Pass | Apache-2.0 | today |
| 105 | Generate vector embeddings via KeiRouter /v1/embeddings using OpenAI / Gemini / Mistral / Voyage / Nvidia embedding models for RAG, semantic search, similarity. | mydisha/ | 147 | — | ~577 | Automated safety check: Pass | MIT | 1 mo ago |
| 106 | Set up a K3s cluster on an NVIDIA GPU host, connect it to Azure Arc, and configure Azure ML to use it as a Kubernetes compute target. | microsoft/ | 126 | — | ~5.7k | Automated safety check: Notes | MIT | 2 days ago |
| 107 | Use MemoryWhale debugging evidence when the user requests recall or a recurring failure may have relevant recorded history. | wuisabel-gif/ | 154 | — | ~459 | Automated safety check: Pass | MIT | 3 days ago |
| 108 | Orchestrates defect image generation for PCBA, metal surface and glass inspection with NVIDIA Cosmos AnomalyGen on OSMO, from cold-start Day 0 to real-photo Day 1 labeling. | NVIDIA/ | 3.6k | — | ~5k | Automated safety check: Notes | Apache-2.0 | 2 days ago |
| 109 | Inventory and explain the patch (monkey-patch) optimizations Primus layers over upstream training backends such as Megatron-LM, TorchTitan, and MaxText, including their version compatibility… | AMD-AGI/ | 131 | — | ~3.5k | Automated safety check: Pass | Unknown | today |
| 110 | Validate the staging NemoClaw Brev Launchable through its web journey or a deployed environment. | NVIDIA/ | 23k | — | ~3.3k | Automated safety check: Pass | Apache-2.0 | today |
| 111 | 111.Cuda Skill Query current NVIDIA CUDA, PTX ISA, Runtime API, Driver API, Programming Guide, Best Practices, Nsight Compute, and Nsight Systems references. | slowlyC/ | 169 | — | ~1.8k | Automated safety check: Pass | MIT | 2 mo ago |
| 112 | Run the Nemotron-3.5 Lightning Text2SQL LoRA fine-tuning tutorial (NeMo Megatron-Bridge) end-to-end for the user on a single node: data prep, checkpoint conversion, LoRA fine-tuning of the 30B-A3B… | NVIDIA-NeMo/ | 2.1k | — | ~1.9k | Automated safety check: Pass | Apache-2.0 | 5 days ago |
| 113 | Orchestrates video data augmentation and auto-labeling workflows on OSMO, from flow selection and preflight checks to submission, monitoring and output download. | NVIDIA/ | 3.6k | — | ~4.7k | Automated safety check: Notes | Apache-2.0 | 2 days ago |
| 114 | 114.Make Op Scaffold Scaffold a new CV-CUDA operator — a complete, wired, building skeleton — and delegate the implementation to a human or another AI. | CVCUDA/ | 2.7k | — | ~306 | Automated safety check: Pass | Unknown | 24 days ago |
| 115 | Calculate training costs for Tinker fine-tuning jobs. An agent skill from sundial-org/skills. | sundial-org/ | 153 | — | ~1.2k | Automated safety check: Pass | No licence | 2 mo ago |
| 116 | Run NVIDIA compute-sanitizer (memcheck, racecheck, initcheck, synccheck) against a Triton/TLX kernel to find runtime memory and synchronization bugs. | facebookexperimental/ | 201 | — | ~1.6k | Automated safety check: Pass | MIT | today |
| 117 | 117.Convergence Test Run, monitor, stop and report Primus convergence tests -- training a model on a real corpus and checking that the loss curve is healthy -- from a plain-language request such as "run convergence test… | AMD-AGI/ | 131 | — | ~2.1k | Automated safety check: Pass | Unknown | today |
| 118 | A skill your agent uses when reviewing the weekly AICR component drift report — the Slack digest and drift-report.json artifact produced by Registry Drift Report (registry-drift.yaml) listing which… | NVIDIA/ | 440 | — | ~2.8k | Automated safety check: Pass | Apache-2.0 | today |
| 119 | Quick install and deploy vLLM, start serving with a simple LLM, and test OpenAI API. | vllm-project/ | 102 | — | ~1.6k | Automated safety check: Pass | Apache-2.0 | 6 mo ago |
| 120 | Optimize fused cross-entropy loss kernels in Triton for NVIDIA and AMD GPUs. | ZJLi2013/ | 102 | — | ~461 | Automated safety check: Pass | No licence | 6 mo ago |
| 121 | Author or edit NeMo Relay documentation or examples when repository-specific MDX, public API, integration, or release-history conventions matter. | NVIDIA/ | 192 | — | ~710 | Automated safety check: Pass | Apache-2.0 | 2 days ago |
| 122 | A skill your agent uses when deploying or operating standalone RTVI-CV-3D / MV3DT multi-camera 3D tracking for calibrated MP4/file inputs and live RTSP streams: missing-calibration handoff to AMC… | NVIDIA-AI-Blueprints/ | 1.9k | — | ~5.1k | Automated safety check: Notes | Apache-2.0 | yesterday |
| 123 | 为 AutoResearch 的双 Agent 轨迹、付费 GPU 长跑、Docker 执行、可信评测与恢复建立共享协议、成本决策和隔离边界。用于小时/包日选择、启动或恢复 campaign、设计证据与防止题目或轨迹串用;不替代具体任务算法或最终平台 QA。 | bosprimigenious/ | 153 | — | ~553 | Automated safety check: Pass | MIT | 6 days ago |
| 124 | Uses Meta's LlamaGuard moderation model to screen prompts and model replies against six safety categories, with vLLM, FastAPI and NeMo Guardrails setups. | Orchestra-Research/ | 13k | 2 repos | ~2.3k | Automated safety check: Pass | MIT | 3 mo ago |
| 125 | Run the Nemotron-3 Ultra Text2SQL LoRA fine-tuning tutorial (NeMo Megatron-Bridge) end-to-end for the user on their SLURM cluster: data prep, distributed checkpoint conversion, and packed LoRA… | NVIDIA-NeMo/ | 2.1k | — | ~1.9k | Automated safety check: Pass | Apache-2.0 | 5 days ago |
| 126 | Runs NVIDIA TAO Data Services KPI analysis on object detection results, comparing predictions to ground truth and writing per-class precision, recall and AP to a CSV. | NVIDIA/ | 3.6k | — | ~2.7k | Automated safety check: Notes | Apache-2.0 | 2 days ago |
| 127 | 127.Module 1 This skill should be used when a learner is working through Module 1 ("Build an Agent") of the Build-an-Agent workshop and wants help understanding the concepts, notebooks, or code — e.g. | brevdev/ | 146 | — | ~2.7k | Automated safety check: Pass | Apache-2.0 | 3 days ago |
| 128 | 128.Make Op Verify Verify a new CV-CUDA operator against the deterministic final regression checklist (the /make-op done-gate). | CVCUDA/ | 2.7k | — | ~433 | Automated safety check: Pass | Unknown | 24 days ago |
| 129 | Collect and normalize environment facts (OS, Python, GPU, CUDA/ROCm, container state) before Quark installation or PTQ planning. | amd/ | 182 | — | ~1.4k | Automated safety check: Pass | MIT | 13 days ago |
| 130 | Split a PR into multiple PRs to reduce the number of required CODEOWNERS reviewer groups. | NVIDIA/ | 18k | — | ~1.4k | Automated safety check: Pass | Apache-2.0 | today |
| 131 | Optimize FlashAttention-style fused attention kernels in Triton for NVIDIA and AMD GPUs. | ZJLi2013/ | 102 | — | ~697 | Automated safety check: Pass | No licence | 6 mo ago |
| 132 | 132.Nemotron Ultra Reference desk for NVIDIA Nemotron 3 Ultra (550B-A55B) — architecture, NVFP4 pretraining, SFT, MOPD (multi-teacher on-policy distillation), MTP boosting, quantization, inference. | NVIDIA-NeMo/ | 2.1k | — | ~1.8k | Automated safety check: Pass | Apache-2.0 | 5 days ago |
| 133 | A skill your agent uses when an SGLang, vLLM, TensorRT-LLM, or TokenSpeed serving/model optimization task needs prior model-family PR evidence. | BBuf/ | 938 | — | ~1.5k | Automated safety check: Pass | No licence | 6 days ago |
| 134 | 134.Upgrade Deps Upgrade focused runtime dependencies in AReaL. An agent skill from areal-project/AReaL. | areal-project/ | 5.8k | — | ~6k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 135 | MemoryWhale debugging; compiler failures; terminal diagnostics. | wuisabel-gif/ | 154 | — | ~370 | Automated safety check: Pass | MIT | 3 days ago |
| 136 | Runs TAO Data Services gap analysis that compares ground-truth and predicted boxes to find weak images by per-class recall, precision and AP50. | NVIDIA/ | 3.6k | — | ~1.8k | Automated safety check: Notes | Apache-2.0 | 2 days ago |
| 137 | 137.Module 1 This skill should be used when a learner is working through Module 1 ("Build an Agent") of the Build-an-Agent workshop and wants help understanding the concepts, notebooks, or code — e.g. | brevdev/ | 146 | — | ~2.7k | Automated safety check: Pass | Apache-2.0 | 3 days ago |
| 138 | Review a CV-CUDA operator's BENCHMARK coverage — drivers, layout axis, baselines, the basic-tier floor, row counts, and coverage statistics. | CVCUDA/ | 2.7k | — | ~274 | Automated safety check: Pass | Unknown | 24 days ago |
| 139 | Classify one failed NemoClaw GitHub Actions job using bounded, redacted logs and optional retained artifacts. | NVIDIA/ | 23k | — | ~806 | Automated safety check: Pass | Apache-2.0 | today |
| 140 | Research and draft a response to a GitHub issue or question from an external contributor. | NVIDIA/ | 18k | — | ~915 | Automated safety check: Pass | Unknown | today |
| 141 | 141.Watchdog Hold / release GPUs by running a standard SGLang inference server on them (lab convention — occupy a card with a real serving job that fills ~all of its memory AND is driven at stable near-full… | cua-lite/ | 108 | — | ~1.9k | Automated safety check: Pass | No licence | 2 days ago |
| 142 | 142.Fused Moe Kernel Optimize Fused Mixture-of-Experts (MoE) kernels in Triton for NVIDIA and AMD GPUs. | ZJLi2013/ | 102 | — | ~699 | Automated safety check: Pass | No licence | 6 mo ago |
| 143 | Sets up and runs NVIDIA Cosmos Policy evaluations on the LIBERO and RoboCasa simulators, including headless GPU rendering and inference latency profiling. | Orchestra-Research/ | 13k | — | ~3.7k | Automated safety check: Pass | MIT | 3 mo ago |
| 144 | Fine-tunes and evaluates OpenVLA-OFT and OFT+ robot policies with LoRA and continuous action heads on LIBERO simulation and ALOHA real-robot setups. | Orchestra-Research/ | 13k | — | ~3.7k | Automated safety check: Pass | MIT | 3 mo ago |