AI model or service
NVIDIA AI Platform agent skills, page 9
NVIDIA AI Platform skills, ranked
Ranked by score. Sort bymost stars,trending,newest,recently updated
| # | Skill | Repository | Stars | Used in | Tokens | Auto-check | Licence | Updated |
|---|---|---|---|---|---|---|---|---|
| 385 | 385.Optimize For GPU GPU-accelerates scientific Python on NVIDIA hardware and verifies that the result is correct and faster. | K-Dense-AI/ | 48k | 1 repo | ~3.4k | Automated safety check: Pass | MIT | 2 days ago |
| 386 | Review a CV-CUDA operator's input-type, layout, dtype, and channel support matrix. | CVCUDA/ | 2.7k | — | ~248 | Automated safety check: Pass | Unknown | 20 days ago |
| 387 | Track and normalize change requests against the official Megatron-LM repository by branch, PR, commit, commit range, or time window. | ascend-ai-coding/ | 174 | — | ~1.1k | Automated safety check: Pass | No licence | today |
| 388 | 388.Model Serving LLM and ML model deployment for inference. An agent skill from ancoleman/ai-design-components. | ancoleman/ | 526 | 1 repo | ~3.4k | Automated safety check: Pass | MIT | 10 mo ago |
| 389 | 389.GPU Status Show GPU availability across all SSH servers listed in this project's CLAUDE.md. | CurryTang/ | 176 | — | ~862 | Automated safety check: Warn | No licence | 6 mo ago |
| 390 | Guidance for Azure Confidential Computing — protecting data in use through hardware-based Trusted Execution Environments (TEEs). | vinayaklatthe/ | 175 | — | ~2.4k | Automated safety check: Pass | MIT | 3 mo ago |
| 391 | Opinionated guide for using Primus Projection to choose parallelism (TP/PP/EP/CP/DP) and pipeline schedules, validate memory fit on target nodes, reason about communication collectives, and explore… | AMD-AGI/ | 130 | — | ~5.3k | Automated safety check: Pass | Unknown | today |
| 392 | Review a CV-CUDA operator's test coverage, including C++ correctness, required cross-layout parity, correctness rigor, and the Python API surface. | CVCUDA/ | 2.7k | — | ~264 | Automated safety check: Pass | Unknown | 20 days ago |
| 393 | Map migration-relevant Megatron changes onto the official MindSpeed repository by resolving branch alignment, locating affected subsystems, and identifying concrete adaptation points. | ascend-ai-coding/ | 174 | — | ~1.3k | Automated safety check: Pass | No licence | today |
| 394 | Install, configure, verify, repair, update, and uninstall Hyprland on Fedora Linux with GPU-aware detection (NVIDIA/AMD/Intel). | sickn33/ | 47k | 1 repo | ~1.8k | Automated safety check: Notes | MIT | yesterday |
| 395 | Operate GPU-backed Kubernetes clusters for AI inference and training with scheduling, autoscaling, node health, MIG partitioning, and cost controls. | sickn33/ | 47k | 2 repos | ~3.2k | Automated safety check: Pass | MIT | yesterday |
| 396 | Set up and manage NVIDIA GPU servers for AI workloads. An agent skill from sickn33/agentic-awesome-skills. | sickn33/ | 47k | 2 repos | ~2k | Automated safety check: Notes | MIT | yesterday |
| 397 | Implements input/output validation guardrails for LLM applications using NVIDIA NeMo Guardrails (Colang), custom Python validators for PII detection, and the Guardrails AI framework, intercepting… | mukul975/ | 34k | — | ~2.2k | Automated safety check: Pass | Apache-2.0 | 1 mo ago |
| 398 | Implements defense-in-depth controls at an AI agent's tool-invocation boundary using tool allowlisting, least-privilege identity binding, NeMo Guardrails policy enforcement, human-in-the-loop… | mukul975/ | 34k | — | ~3k | Automated safety check: Pass | Apache-2.0 | 1 mo ago |
| 399 | Probes Retrieval-Augmented Generation pipelines for indirect prompt injection via poisoned retrieved documents and embedding-space manipulation, using NVIDIA garak, Promptfoo red-team plugins, and… | mukul975/ | 34k | — | ~3.3k | Automated safety check: Pass | Apache-2.0 | 1 mo ago |
| 400 | 400.Kernel Profiling Profile GPU kernels using NCU (NVIDIA) or rocprof (AMD) to collect performance metrics. | ZJLi2013/ | 102 | — | ~696 | Automated safety check: Pass | No licence | 6 mo ago |
| 401 | Set up a working ML training/inference environment on NVIDIA DGX Spark (GB10, aarch64, CUDA 13). | wshobson/ | 40k | — | ~2k | Automated safety check: Pass | MIT | 2 days ago |
| 402 | Deploy ML models on Kubernetes with KServe (formerly KFServing) and NVIDIA Triton Inference Server. | sickn33/ | 47k | 1 repo | ~2.3k | Automated safety check: Pass | MIT | yesterday |
| 403 | Check whether a CV-CUDA operator is READY to optimize (correctness + bench coverage + captured baseline + profiling) per .agents/guidance/OPTIMIZATIONGUIDELINES.md. | CVCUDA/ | 2.7k | — | ~255 | Automated safety check: Pass | Unknown | 20 days ago |
| 404 | 404.Mangohud Overlay Diagnose or change how the MangoHud HUD and CSV log attach to the game across launchers (Steam, direct), renderers (NVIDIA GL, Zink) and display servers (X11, native Wayland). | xD3I/ | 113 | — | ~493 | Automated safety check: Pass | No licence | today |
| 405 | Optimize RMS Normalization kernels in Triton for NVIDIA and AMD GPUs. | ZJLi2013/ | 102 | — | ~728 | Automated safety check: Pass | No licence | 6 mo ago |
| 406 | 406.System Profile Profile a target (script, process, GPU, memory, interconnect) using external tools and code instrumentation. | AI4Scientist/ | 128 | 4 repos | ~1.1k | Automated safety check: Pass | No licence | 4 mo ago |
| 407 | 407.Accelerate Run PyTorch training across GPUs with minimal changes. An agent skill from Luciole-Studio/Misaka-Agent. | Luciole-Studio/ | 125 | 1 repo | ~2.3k | Automated safety check: Pass | MIT | today |
| 408 | 408.Slime RL post-training for LLMs with Megatron and SGLang. An agent skill from Luciole-Studio/Misaka-Agent. | Luciole-Studio/ | 125 | 1 repo | ~2.7k | Automated safety check: Pass | MIT | today |
| 409 | 409.Tensorrt LLM High-throughput LLM inference on NVIDIA GPUs. An agent skill from Luciole-Studio/Misaka-Agent. | Luciole-Studio/ | 125 | 1 repo | ~1.3k | Automated safety check: Pass | MIT | today |
| 410 | Automatically fetch InferenceX benchmark data and generate daily performance reports for LLM inference on various hardware (NVIDIA, AMD, etc.). | ascend-ai-coding/ | 174 | — | ~1.1k | Automated safety check: Pass | No licence | today |
| 411 | Launch / relaunch agentic RL (SkyRL terminalbench + Harbor + Daytona) on JSC Jupiter (GH200). | open-thoughts/ | 301 | — | ~2.7k | Automated safety check: Pass | Apache-2.0 | 9 days ago |
| 412 | Optimize fused softmax kernels in Triton for NVIDIA and AMD GPUs. | ZJLi2013/ | 102 | — | ~742 | Automated safety check: Pass | No licence | 6 mo ago |
| 413 | Runs NVIDIA garak probe suites (jailbreak, prompt injection, data leakage, toxicity, and more) against an LLM endpoint - Hugging Face models, OpenAI-compatible APIs, or Bedrock - then interprets the… | mukul975/ | 34k | — | ~2.9k | Automated safety check: Warn | Apache-2.0 | 1 mo ago |
| 414 | Cross-engine decision rubric for self-hosting or recommending an LLM serving stack. | agentsope/ | 457 | — | ~6.1k | Automated safety check: Pass | MIT | 1 mo ago |
| 415 | 415.Agentsop Vllm Decision SOP for serving LLMs with vLLM. An agent skill from agentsope/SkillAlchemy. | agentsope/ | 457 | — | ~6.1k | Automated safety check: Pass | MIT | 1 mo ago |
| 416 | 416.Engine Sglang Serve a Hugging Face model with SGLang on a Linux machine with an NVIDIA or AMD GPU, configured from the model's SGLang cookbook page — or, when it has none, from the model's own files — and join it… | autonomous-ai/ | 1.1k | — | ~1.5k | Automated safety check: Pass | MIT | today |
| 417 | 417.Engine Vllm Serve a Hugging Face model with vLLM on a Linux machine with an NVIDIA or AMD GPU, configured from the model's official vLLM recipe — or, when it has none, from the model's own files — and join it… | autonomous-ai/ | 1.1k | — | ~1.9k | Automated safety check: Pass | MIT | today |
| 418 | 418.Nvwb Project Apply when working in a directory containing .project/spec.yaml — provides NVIDIA AI Workbench project awareness for in-container development. | brevdev/ | 143 | — | ~1.5k | Automated safety check: Warn | Apache-2.0 | today |
| 419 | Computer vision engineering for object detection, segmentation, and visual AI, covering CNN and Vision Transformer architectures and ONNX/TensorRT deployment. | borghei/ | 874 | — | ~1.8k | Automated safety check: Pass | MIT | today |
| 420 | 420.Nemo A skill your agent uses for NVIDIA NeMo Speech ASR, TTS, audio processing, SpeechLM2 and voice-agent workflows, speaker diarization, data tooling, and repository development. | majiayu000/ | 666 | 1 repo | ~1.5k | Automated safety check: Pass | Apache-2.0 | today |
| 421 | Optimize Rotary Position Embedding (RoPE) kernels in Triton for NVIDIA and AMD GPUs. | ZJLi2013/ | 102 | — | ~421 | Automated safety check: Pass | No licence | 6 mo ago |
| 422 | 422.Serving Systems LLM and multimodal serving systems. An agent skill from uw-syfi/vibesys. | uw-syfi/ | 103 | — | ~2.9k | Automated safety check: Pass | MIT | yesterday |
| 423 | Run GPU jobs on NVIDIA NIM microservices via host.compute.create('byoc:nvidia', ...). | PKU-YuanGroup/ | 608 | — | ~2.9k | Automated safety check: Pass | Apache-2.0 | 2 days ago |
| 424 | Verl 分布式训练服务一键拉起与配置。触发场景:(1) 用户要启动 Verl 训练任务或部署 RLHF/DAPO 训练环境 (2) 在 NPU 集群上拉起 Verl 训练容器 (3) 配置 Ray 集群和 SwanLab 监控 (4) 根据 7 位二进制掩码灵活配置加速特性。支持 Qwen3-8B 等 Megatron 模型的 DAPO 训练全流程。 | ascend-ai-coding/ | 174 | — | ~2k | Automated safety check: Pass | No licence | today |
| 425 | 425.Optimize For GPU GPU-accelerate Python code using CuPy, Numba CUDA, Warp, cuDF, cuML, cuGraph, KvikIO, cuCIM, cuxfilter, cuVS, cuSpatial, and RAFT. | majiayu000/ | 666 | 1 repo | ~8.5k | Automated safety check: Pass | MIT | today |
| 426 | Structured, multi-dimensional company investment research framework for AI agents and human analysts. | aAAaqwq/ | 105 | 1 repo | ~4k | Automated safety check: Pass | MIT | 10 days ago |
| 427 | 427.GPU Optimizer GPU optimization for consumer NVIDIA GPUs (8-24GB VRAM) covering mixed precision, gradient checkpointing, XGBoost GPU, CuPy/cuDF migration, and torch.compile. | Mathews-Tom/ | 327 | — | ~3.5k | Automated safety check: Notes | MIT | yesterday |
| 428 | Use this sub-skill for Torch-TensorRT model compilation, dynamic input planning, torch.export workflows, save/load formats, raw TensorRT engines, and compile-time troubleshooting. | VectorSpaceLab/ | 328 | — | ~1.3k | Automated safety check: Pass | BSD-3-Clause | 1 mo ago |
| 429 | Use this sub-skill for Torch-TensorRT runtime performance controls, CUDA Graphs, output allocation, caches, TensorRT-RTX runtime settings, mutable modules, refit, weight streaming, and benchmark… | VectorSpaceLab/ | 328 | — | ~1k | Automated safety check: Pass | BSD-3-Clause | 1 mo ago |
| 430 | 430.Torch Tensorrt A skill your agent uses for Torch-TensorRT tasks: compiling PyTorch models with TensorRT, dynamic-shape/export workflows, runtime optimization, Triton/C++/distributed deployment, debugging… | VectorSpaceLab/ | 328 | — | ~1.5k | Automated safety check: Pass | BSD-3-Clause | 1 mo ago |
| 431 | 431.Setup Prover Set up and deploy a Boundless prover to a GPU server using Ansible. | boundless-xyz/ | 193 | — | ~4.1k | Automated safety check: Warn | Apache-2.0 | 1 mo ago |
| 432 | Generate migration deliverables for bringing relevant Megatron changes into MindSpeed after branch alignment and impact mapping are complete. | ascend-ai-coding/ | 174 | — | ~1.3k | Automated safety check: Pass | No licence | today |