Search
NVIDIA AI Platform
Skills
Sort:BestMost starsTrending todayTrending this weekTrending this monthNewestRecently updatedName
| # | Skill | Repository | Stars | Used in | Tokens | Auto-check | Licence | Updated |
|---|---|---|---|---|---|---|---|---|
| 433 | Structured, multi-dimensional company investment research framework for AI agents and human analysts. | aAAaqwq/ | 105 | 1 repo | ~4k | Automated safety check: Pass | MIT | 3 days ago |
| 434 | 434.GPU Optimizer GPU optimization for consumer NVIDIA GPUs (8-24GB VRAM) covering mixed precision, gradient checkpointing, XGBoost GPU, CuPy/cuDF migration, and torch.compile. | Mathews-Tom/ | 329 | — | ~3.5k | Automated safety check: Notes | MIT | 5 days ago |
| 435 | Use this sub-skill for Torch-TensorRT model compilation, dynamic input planning, torch.export workflows, save/load formats, raw TensorRT engines, and compile-time troubleshooting. | VectorSpaceLab/ | 331 | — | ~1.3k | Automated safety check: Pass | BSD-3-Clause | 1 mo ago |
| 436 | Use this sub-skill for Torch-TensorRT runtime performance controls, CUDA Graphs, output allocation, caches, TensorRT-RTX runtime settings, mutable modules, refit, weight streaming, and benchmark… | VectorSpaceLab/ | 331 | — | ~1k | Automated safety check: Pass | BSD-3-Clause | 1 mo ago |
| 437 | 437.Torch Tensorrt A skill your agent uses for Torch-TensorRT tasks: compiling PyTorch models with TensorRT, dynamic-shape/export workflows, runtime optimization, Triton/C++/distributed deployment, debugging… | VectorSpaceLab/ | 331 | — | ~1.5k | Automated safety check: Pass | BSD-3-Clause | 1 mo ago |
| 438 | 438.Setup Prover Set up and deploy a Boundless prover to a GPU server using Ansible. | boundless-xyz/ | 193 | — | ~4.1k | Automated safety check: Warn | Apache-2.0 | 1 mo ago |
| 439 | Generate migration deliverables for bringing relevant Megatron changes into MindSpeed after branch alignment and impact mapping are complete. | ascend-ai-coding/ | 174 | — | ~1.3k | Automated safety check: Pass | No licence | yesterday |
| 440 | Verl 单异步 DAPO 训练配置生成器。触发场景:(1) 启动单异步 DAPO 训练 (2) 生成训练脚本 (3) 配置特性参数 (4) 训练前检查。特性策略:用户未指定时默认开启性能特性(flashattn/dynamicbatch/removepadding/gradientcheckpointing),显存特性(offload/recompute)默认关闭。OOM… | ascend-ai-coding/ | 174 | — | ~1.2k | Automated safety check: Pass | No licence | yesterday |
| 441 | MindSpeed-MM multimodal model suite environment setup guide for Huawei Ascend NPU. | ascend-ai-coding/ | 174 | — | ~3.1k | Automated safety check: Pass | No licence | yesterday |
| 442 | 442.Mindspeed Mm Vlm Universal VLM (vision-language understanding model) training guide for Huawei Ascend NPU using MindSpeed-MM. | ascend-ai-coding/ | 174 | — | ~5.2k | Automated safety check: Pass | No licence | yesterday |
| 443 | 443.Verl Quickstart Generates an executable, end-to-end VERL reinforcement learning quickstart runbook for Ascend/NPU (docker image, dataset preprocessing, model setup, mainppo training, and examples/run.sh flow). | ascend-ai-coding/ | 174 | — | ~592 | Automated safety check: Pass | No licence | yesterday |
| 444 | One-time session setup and orchestration map for the TAO skill bank. | NVIDIA/ | 3.6k | — | ~1.8k | Automated safety check: Warn | Apache-2.0 | 2 days ago |
| 445 | Linting and formatting for Megatron-LM. An agent skill from NVIDIA/skills. | NVIDIA/ | 3.6k | — | ~406 | Automated safety check: Pass | Apache-2.0 | 2 days ago |
| 446 | 446.Parakeet Stt Local speech-to-text with NVIDIA Parakeet TDT 0.6B v3 (ONNX on CPU). | sundial-org/ | 663 | — | ~771 | Automated safety check: Pass | No licence | 7 mo ago |
| 447 | 447.Cuda CUDA C/C++ skill for NVIDIA GPU kernel programming. An agent skill from mohitmishra786/low-level-dev-skills. | mohitmishra786/ | 252 | — | ~1.9k | Automated safety check: Pass | MIT | 3 mo ago |
| 448 | 448.Cuda Debugging CUDA debugging skill for GPU program correctness. An agent skill from mohitmishra786/low-level-dev-skills. | mohitmishra786/ | 252 | — | ~1.5k | Automated safety check: Pass | MIT | 3 mo ago |
| 449 | 449.Cuda Profiling CUDA profiling skill for NVIDIA GPU performance analysis. An agent skill from mohitmishra786/low-level-dev-skills. | mohitmishra786/ | 252 | — | ~1.6k | Automated safety check: Notes | MIT | 3 mo ago |
| 450 | 450.GPU Memory Model GPU memory model skill for SIMT execution and memory hierarchy. | mohitmishra786/ | 252 | — | ~1.9k | Automated safety check: Pass | MIT | 3 mo ago |
| 451 | Proteina-Complexa flow-based protein backbone generation with fold-conditioned sampling guidance. | BioTender-max/ | 200 | — | ~1.4k | Automated safety check: Notes | MIT | 3 mo ago |
| 452 | Build machine vision inspection systems with MATLAB Visual Inspection Toolbox. | matlab/ | 1.1k | — | ~3.1k | Automated safety check: Pass | Unknown | 3 days ago |
| 453 | Bind pre-downloaded Jetson reference docs (developer guide, design guide, pinmux, schematics) into the active profile documents block. | NVIDIA/ | 3.6k | — | ~3.3k | Automated safety check: Pass | Apache-2.0 | 2 days ago |
| 454 | Audit, prepare, and deploy PAIDF Orchestration on a Kubernetes GPU cluster - single-GPU H100/L40S hosts, managed Kubernetes, kubeadm, and similar. | NVIDIA/ | 3.6k | — | ~3.8k | Automated safety check: Warn | Apache-2.0 | 2 days ago |
| 455 | The data-mover for TAO jobs — decides the storage tier (A pre-positioned mount with zero fetch / B volume-from-S3 / C ephemeral in-compute fetch), stages inputs (bulk + annotation-selective +… | NVIDIA/ | 3.6k | — | ~1.5k | Automated safety check: Warn | Apache-2.0 | 2 days ago |
| 456 | The Docker execution platform for TAO jobs — a local daemon or a remote GPU box via DOCKERHOST=ssh://user@host. | NVIDIA/ | 3.6k | — | ~5k | Automated safety check: Warn | Apache-2.0 | 2 days ago |
| 457 | Split a PR into multiple PRs to reduce the number of required CODEOWNERS reviewer groups. | NVIDIA/ | 3.6k | — | ~1.1k | Automated safety check: Pass | Apache-2.0 | 2 days ago |