AI model or service
NVIDIA AI Platform agent skills, page 8
NVIDIA AI Platform skills, ranked
Ranked by score. Sort bymost stars,trending,newest,recently updated
| # | Skill | Repository | Stars | Used in | Tokens | Auto-check | Licence | Updated |
|---|---|---|---|---|---|---|---|---|
| 337 | Best practices for Docker-based ROS2 development including multi-stage Dockerfiles, docker-compose for multi-container robotic systems, DDS discovery across containers, GPU passthrough for… | arpitg1304/ | 368 | — | ~9.1k | Automated safety check: Notes | Apache-2.0 | 1 mo ago |
| 338 | Explains where HOT-Step generation time goes (LM/DiT/VAE), how the TensorRT paths activate, how to benchmark from logs, and which knobs trade quality for speed. | scragnog/ | 170 | — | ~4.9k | Automated safety check: Pass | MIT | 2 days ago |
| 339 | Bilingual guide for running and interpreting LLaVA-OneVision2 HF vs Megatron consistency checks across TP and PP settings | EvolvingLMMs-Lab/ | 1.2k | — | ~4.1k | Automated safety check: Pass | Apache-2.0 | today |
| 340 | A skill your agent uses when asked to install, deploy, run, validate, troubleshoot, or stop NVIDIA Deep Researcher Agent Blueprint infrastructure. | NVIDIA-AI-Blueprints/ | 883 | — | ~3.5k | Automated safety check: Notes | Apache-2.0 | yesterday |
| 341 | A skill your agent uses when accelerating existing genomics workflows with NVIDIA Parabricks, improving runtime or price/performance, converting pipeline steps to GPUs, or comparing CPU and GPU… | NVIDIA-BioNeMo/ | 478 | — | ~3.3k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 342 | Estimate GPU memory usage for Megatron-based MoE (Mixture of Experts) and dense models. | yzlnew/ | 149 | — | ~2.2k | Automated safety check: Pass | No licence | 3 mo ago |
| 343 | Detects CPU, GPU, memory and disk resources before heavy scientific tasks and writes a JSON file with advice on parallelism, out-of-core work and GPU use. | davila7/ | 32k | 10 repos | ~2.4k | Automated safety check: Pass | MIT | today |
| 344 | 344.Init GPU Server Initialize a Draw Things GPU server with GPUScript, including script sync, Docker/CUDA/NVIDIA runtime setup, 7T data disk mounting, mergerfs, and end-to-end GPU verification. | drawthingsai/ | 579 | — | ~2.2k | Automated safety check: Pass | GPL-3.0 | yesterday |
| 345 | 345.Nvwb Manage NVIDIA AI Workbench projects, contexts, builds, and environments via the nvwb CLI. | brevdev/ | 143 | — | ~1.1k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 346 | 346.Fla Mr Readiness Checklist and workflow for preparing an MR/PR in the FLA repo. | fla-org/ | 5.8k | — | ~1.3k | Automated safety check: Pass | MIT | yesterday |
| 347 | Generate vector embeddings via KeiRouter /v1/embeddings using OpenAI / Gemini / Mistral / Voyage / Nvidia embedding models for RAG, semantic search, similarity. | mydisha/ | 147 | — | ~577 | Automated safety check: Pass | MIT | 26 days ago |
| 348 | Use MemoryWhale debugging evidence when the user requests recall or a recurring failure may have relevant recorded history. | wuisabel-gif/ | 151 | — | ~459 | Automated safety check: Pass | MIT | yesterday |
| 349 | Inventory and explain the patch (monkey-patch) optimizations Primus layers over upstream training backends such as Megatron-LM, TorchTitan, and MaxText, including their version compatibility… | AMD-AGI/ | 130 | — | ~3.5k | Automated safety check: Pass | Unknown | today |
| 350 | 350.Cuda Skill Query current NVIDIA CUDA, PTX ISA, Runtime API, Driver API, Programming Guide, Best Practices, Nsight Compute, and Nsight Systems references. | slowlyC/ | 169 | — | ~1.8k | Automated safety check: Pass | MIT | 2 mo ago |
| 351 | A skill your agent uses when deploying or operating standalone RTVI-CV-3D / MV3DT multi-camera 3D tracking for calibrated MP4/file inputs and live RTSP streams: missing-calibration handoff to AMC… | NVIDIA-AI-Blueprints/ | 1.9k | — | ~5.1k | Automated safety check: Notes | Apache-2.0 | today |
| 352 | Run the Nemotron-3.5 Lightning Text2SQL LoRA fine-tuning tutorial (NeMo Megatron-Bridge) end-to-end for the user on a single node: data prep, checkpoint conversion, LoRA fine-tuning of the 30B-A3B… | NVIDIA-NeMo/ | 2.1k | — | ~1.9k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 353 | Uses Meta's LlamaGuard moderation model to screen prompts and model replies against six safety categories, with vLLM, FastAPI and NeMo Guardrails setups. | Orchestra-Research/ | 13k | 3 repos | ~2.3k | Automated safety check: Pass | MIT | 3 mo ago |
| 354 | Guidelines for NVIDIA GPU kernel / Triton / Gluon / TileLang / CUDA backend performance work in the FLA repo. | fla-org/ | 5.8k | — | ~1k | Automated safety check: Pass | MIT | yesterday |
| 355 | 355.Make Op Scaffold Scaffold a new CV-CUDA operator — a complete, wired, building skeleton — and delegate the implementation to a human or another AI. | CVCUDA/ | 2.7k | — | ~306 | Automated safety check: Pass | Unknown | 21 days ago |
| 356 | Calculate training costs for Tinker fine-tuning jobs. An agent skill from sundial-org/skills. | sundial-org/ | 152 | — | ~1.2k | Automated safety check: Pass | No licence | 2 mo ago |
| 357 | Fine-tunes and evaluates OpenVLA-OFT and OFT+ robot policies with LoRA and continuous action heads on LIBERO simulation and ALOHA real-robot setups. | Orchestra-Research/ | 13k | 1 repo | ~3.7k | Automated safety check: Pass | MIT | 3 mo ago |
| 358 | Quick install and deploy vLLM, start serving with a simple LLM, and test OpenAI API. | vllm-project/ | 103 | — | ~1.6k | Automated safety check: Pass | Apache-2.0 | 6 mo ago |
| 359 | Optimize fused cross-entropy loss kernels in Triton for NVIDIA and AMD GPUs. | ZJLi2013/ | 102 | — | ~461 | Automated safety check: Pass | No licence | 6 mo ago |
| 360 | 为 AutoResearch 的双 Agent 轨迹、付费 GPU 长跑、Docker 执行、可信评测与恢复建立共享协议、成本决策和隔离边界。用于小时/包日选择、启动或恢复 campaign、设计证据与防止题目或轨迹串用;不替代具体任务算法或最终平台 QA。 | bosprimigenious/ | 149 | — | ~553 | Automated safety check: Pass | MIT | 3 days ago |
| 361 | Run the Nemotron-3 Ultra Text2SQL LoRA fine-tuning tutorial (NeMo Megatron-Bridge) end-to-end for the user on their SLURM cluster: data prep, distributed checkpoint conversion, and packed LoRA… | NVIDIA-NeMo/ | 2.1k | — | ~1.9k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 362 | 362.Make Op Verify Verify a new CV-CUDA operator against the deterministic final regression checklist (the /make-op done-gate). | CVCUDA/ | 2.7k | — | ~433 | Automated safety check: Pass | Unknown | 21 days ago |
| 363 | 363.Module 1 This skill should be used when a learner is working through Module 1 ("Build an Agent") of the Build-an-Agent workshop and wants help understanding the concepts, notebooks, or code — e.g. | brevdev/ | 143 | — | ~2.7k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 364 | Collect and normalize environment facts (OS, Python, GPU, CUDA/ROCm, container state) before Quark installation or PTQ planning. | amd/ | 181 | — | ~1.4k | Automated safety check: Pass | MIT | 9 days ago |
| 365 | Optimize FlashAttention-style fused attention kernels in Triton for NVIDIA and AMD GPUs. | ZJLi2013/ | 102 | — | ~697 | Automated safety check: Pass | No licence | 6 mo ago |
| 366 | 366.Nemotron Ultra Reference desk for NVIDIA Nemotron 3 Ultra (550B-A55B) — architecture, NVFP4 pretraining, SFT, MOPD (multi-teacher on-policy distillation), MTP boosting, quantization, inference. | NVIDIA-NeMo/ | 2.1k | — | ~1.8k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 367 | A skill your agent uses when an SGLang, vLLM, TensorRT-LLM, or TokenSpeed serving/model optimization task needs prior model-family PR evidence. | BBuf/ | 900 | — | ~1.5k | Automated safety check: Pass | No licence | 2 days ago |
| 368 | 368.Upgrade Deps Upgrade focused runtime dependencies in AReaL. An agent skill from areal-project/AReaL. | areal-project/ | 5.8k | — | ~6k | Automated safety check: Pass | Apache-2.0 | 4 days ago |
| 369 | Review a CV-CUDA operator's BENCHMARK coverage — drivers, layout axis, baselines, the basic-tier floor, row counts, and coverage statistics. | CVCUDA/ | 2.7k | — | ~274 | Automated safety check: Pass | Unknown | 21 days ago |
| 370 | MemoryWhale debugging; compiler failures; terminal diagnostics. | wuisabel-gif/ | 151 | — | ~370 | Automated safety check: Pass | MIT | yesterday |
| 371 | 371.Module 1 This skill should be used when a learner is working through Module 1 ("Build an Agent") of the Build-an-Agent workshop and wants help understanding the concepts, notebooks, or code — e.g. | brevdev/ | 143 | — | ~2.7k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 372 | 372.Watchdog Hold / release GPUs by running a standard SGLang inference server on them (lab convention — occupy a card with a real serving job that fills ~all of its memory AND is driven at stable near-full… | cua-lite/ | 106 | — | ~1.9k | Automated safety check: Pass | No licence | 11 days ago |
| 373 | 373.Fused Moe Kernel Optimize Fused Mixture-of-Experts (MoE) kernels in Triton for NVIDIA and AMD GPUs. | ZJLi2013/ | 102 | — | ~699 | Automated safety check: Pass | No licence | 6 mo ago |
| 374 | Sets up and runs NVIDIA Cosmos Policy evaluations on the LIBERO and RoboCasa simulators, including headless GPU rendering and inference latency profiling. | Orchestra-Research/ | 13k | — | ~3.7k | Automated safety check: Pass | MIT | 3 mo ago |
| 375 | Bilingual guidance for Megatron checkpoint 1D 2D 3D mprank layouts across tensor pipeline and expert parallel dimensions | EvolvingLMMs-Lab/ | 1.2k | — | ~1k | Automated safety check: Pass | Apache-2.0 | today |
| 376 | Computer vision engineering skill for object detection, image segmentation, and visual AI systems. | alirezarezvani/ | 28k | 2 repos | ~3.2k | Automated safety check: Pass | MIT | 1 mo ago |
| 377 | 377.Slime User Guide for using SLIME (LLM post-training framework for RL Scaling). | yzlnew/ | 149 | — | ~3.2k | Automated safety check: Pass | No licence | 3 mo ago |
| 378 | Review a CV-CUDA operator's DOCS & API artifacts — operatorlist row, Python autofunction (fn + into), Limitations-table-vs-code consistency, docstrings, and SPDX headers. | CVCUDA/ | 2.7k | — | ~270 | Automated safety check: Pass | Unknown | 21 days ago |
| 379 | Analyze official Megatron-LM commits, PRs, and branch change sets to identify feature evolution, candidate breaking changes, and migration-relevant events. | ascend-ai-coding/ | 174 | — | ~1.2k | Automated safety check: Pass | No licence | today |
| 380 | Optimize dense matrix multiplication (GEMM) kernels in Triton for NVIDIA and AMD GPUs. | ZJLi2013/ | 102 | — | ~1.1k | Automated safety check: Pass | No licence | 6 mo ago |
| 381 | Deploy vLLM using Docker (pre-built images or build-from-source) with NVIDIA GPU support and run the OpenAI-compatible server. | vllm-project/ | 103 | — | ~2.5k | Automated safety check: Notes | Apache-2.0 | 6 mo ago |
| 382 | Generate a validation plan and test scaffolding to check a Primus optimization that has been ported into a user's own training framework, across correctness (numerical accuracy vs a reference)… | AMD-AGI/ | 130 | — | ~1.5k | Automated safety check: Pass | Unknown | today |
| 383 | Write, optimize, and debug high-performance AI compute kernels using TileLang (a Python DSL for GPU programming). | yzlnew/ | 149 | — | ~2.4k | Automated safety check: Pass | No licence | 3 mo ago |
| 384 | 384.Nemo Guardrails NVIDIA's runtime safety framework for LLM applications. An agent skill from Orchestra-Research/AI-Research-SKILLs. | Orchestra-Research/ | 13k | 2 repos | ~1.9k | Automated safety check: Warn | MIT | 3 mo ago |