Search

AI & LLM Engineering · NVIDIA AI Platform

246 skills found, page 5.
Search results
#SkillRepositoryStarsUsed inTokensAuto-checkLicenceUpdated
193

Stage 3 of Clinical ASR Flywheel. An agent skill from NVIDIA/skills.

NVIDIA/skills3.6k—~4.6kAutomated safety check: PassApache-2.02 days ago
194

Enable MIPI/GMSL camera sensors on a Jetson Thor or Orin custom carrier by rendering a kernel-DT overlay from the in-tree sensor DTSI.

NVIDIA/skills3.6k—~2.4kAutomated safety check: PassApache-2.02 days ago
195

Per-controller PCIe enable / disable / lanes / link-speed for a Jetson Thor or Orin custom carrier via ODMDATA + kernel-DT overlay.

NVIDIA/skills3.6k—~2.3kAutomated safety check: PassApache-2.02 days ago
196

Configure Jetson UPHY lane allocation (uphy0/uphy1-config) on Orin/Thor custom carriers.

NVIDIA/skills3.6k—~2.4kAutomated safety check: PassApache-2.02 days ago
197

Enable/disable Jetson USB2/USB3 SS ports via kernel-DT overlay.

NVIDIA/skills3.6k—~2.1kAutomated safety check: PassApache-2.02 days ago
198
198.Jetson Init SourceOfficial

Set up the BSP source workspace: LinuxforTegra overlay tracker, bspsources, Crosstool-NG toolchain.

NVIDIA/skills3.6k—~4.3kAutomated safety check: PassApache-2.02 days ago
199
199.Jetson Init TargetOfficial

Author a new Jetson target-platform profile (referencedevkit + optional customcarrier) and update the active pointer.

NVIDIA/skills3.6k—~4.8kAutomated safety check: PassApache-2.02 days ago
200

Convert single-node scripts to multi-node Slurm sbatch jobs and debug common multi-node failures.

NVIDIA/skills3.6k—~3.9kAutomated safety check: PassApache-2.02 days ago
201

Create, refine, or fix NVIDIA voice agents (Cascaded or Omni) with Pipecat or LiveKit.

NVIDIA/skills3.6k—~844Automated safety check: PassApache-2.02 days ago
202
202.RAG BlueprintOfficial

NVIDIA RAG Blueprint — deploy, configure, troubleshoot, and manage.

NVIDIA/skills3.6k—~2.8kAutomated safety check: NotesApache-2.02 days ago
203
203.RAG EvalOfficial

Filesystem RAG benchmarks: corpus/, train.json, evaluaterag.py (RAGAS quality).

NVIDIA/skills3.6k—~2.3kAutomated safety check: NotesApache-2.02 days ago
204
204.RAG PerfOfficial

Performance benchmarking for a deployed NVIDIA RAG Blueprint server: profiling pass + aiperf load test driven by a single YAML config.

NVIDIA/skills3.6k—~4.1kAutomated safety check: PassApache-2.02 days ago
205

How to swap the DeepStream CV detection model in the VSS Alerts Blueprint verification (2dcv) mode - covers ONNX export, custom bbox parsers, compose mount gotchas, nvinfer config, runtime TRT…

NVIDIA/skills3.6k—~4.5kAutomated safety check: NotesApache-2.02 days ago
206

How to swap the VLM in the VSS Alerts Blueprint — covers RTVI-VLM microservice deployment methods, all three VLM consumers (rtvi-vlm, vlm-as-verifier, vss-agent), and health checks.

NVIDIA/skills3.6k—~5kAutomated safety check: NotesApache-2.02 days ago
207

Integrate a HuggingFace Computer Vision model into the NVIDIA TAO Toolkit ecosystem (tao-core config, tao-pytorch trainer, tao-deploy TensorRT pipeline).

NVIDIA/skills3.6k—~4.5kAutomated safety check: NotesApache-2.02 days ago
208
208.Oob Perf AnalysisOfficial

Generate and analyze T1/T2/R roofline reports for PyTorch OOB workloads comparing Intel XPU and NVIDIA CUDA.

intel/torch-xpu-ops115—~681Automated safety check: PassApache-2.0yesterday
209

Configure auto-configure Ollama when user needs local LLM deployment, free AI alternatives, or wants to eliminate hosted API costs.

jeremylongshore/tons-of-skills-marketplace2.8k—~1.4kAutomated safety check: NotesMITyesterday
210

Computer vision engineering for object detection, segmentation, and visual AI, covering CNN and Vision Transformer architectures and ONNX/TensorRT deployment.

borghei/Claude-Skills891—~1.8kAutomated safety check: PassMIT4 days ago
211

Optimize Rotary Position Embedding (RoPE) kernels in Triton for NVIDIA and AMD GPUs.

ZJLi2013/awesome-kernel-skills102—~421Automated safety check: PassNo licence6 mo ago
212

优化实际模型推理链路,将正确性对齐、分段 profiling、显存与数据搬运、TensorRT/ONNX/PyTorch 后端、attention/kernel、FP8/compile、缓存与少步采样、质量回归、GPU 成本和服务验收串成同一实验闭环。当用户要求推理提速、降低显存或 GPU 成本、复现模型效果、定位 GPU 利用率低、优化图像/视频/扩散模型或自托管 LLM 时使用,提供…

majiayu000/spellbook287—~1.1kAutomated safety check: PassMIT3 days ago
213

Bootstrap a custom carrier board by forking carrier files and scaffolding a DT overlay from the reference devkit.

NVIDIA/skills3.6k—~4.2kAutomated safety check: PassApache-2.02 days ago
214
214.Jetson Generate KbOfficial

Build a per-target knowledge-base markdown next to the active profile by walking the BSP root and source tree.

NVIDIA/skills3.6k—~3.8kAutomated safety check: PassApache-2.02 days ago
215
215.Jetson Init ImageOfficial

Extract Jetson Linux + sample-rootfs tarballs and run applybinaries.sh for the active target, then record bspimage in the profile.

NVIDIA/skills3.6k—~1.8kAutomated safety check: NotesApache-2.02 days ago
216
216.Jetson Quick StartOfficial

Entry skill for Jetson / IGX BSP customization. An agent skill from NVIDIA/skills.

NVIDIA/skills3.6k—~4.7kAutomated safety check: PassApache-2.02 days ago
217
217.Jetson Set TargetOfficial

Switch the active Jetson target-platform pointer to an existing profile YAML.

NVIDIA/skills3.6k—~1.7kAutomated safety check: PassApache-2.02 days ago
218

Guide for selecting and configuring distributed training strategies in NeMo AutoModel, including FSDP2, Megatron FSDP, DDP, and parallelism settings.

NVIDIA/skills3.6k—~4.8kAutomated safety check: PassApache-2.02 days ago
219

Run Megatron-LM (MLM) and Megatron Bridge training with mock or real data.

NVIDIA/skills3.6k—~1.6kAutomated safety check: PassApache-2.02 days ago
220

Validate and use CPU offloading in Megatron Bridge, including layer-level activation offloading and fractional optimizer state offloading with HybridDeviceOptimizer.

NVIDIA/skills3.6k—~2.3kAutomated safety check: PassApache-2.02 days ago
221

Validate and use CUDA graph capture in Megatron Bridge, including local full-iteration graphs and Transformer Engine scoped graphs for attention, MLP, and MoE modules.

NVIDIA/skills3.6k—~3.5kAutomated safety check: PassApache-2.02 days ago
222

Validate and use MoE expert-parallel communication overlap in Megatron-Bridge, including overlapmoeexpertparallelcomm, delaywgradcompute, and flex dispatcher backends such as DeepEP and HybridEP.

NVIDIA/skills3.6k—~3.5kAutomated safety check: PassApache-2.02 days ago
223

Operational guide for enabling Megatron FSDP in Megatron-Bridge, including config knobs, code anchors, pitfalls, and verification.

NVIDIA/skills3.6k—~973Automated safety check: PassApache-2.02 days ago
224

Techniques for reducing peak GPU memory in Megatron Bridge — expandable segments, PEFT + SP input re-gather, parallelism resizing, activation recompute, CPU offloading constraints, and common OOM…

NVIDIA/skills3.6k—~3.6kAutomated safety check: PassApache-2.02 days ago
225

MoE expert-parallel communication overlap in Megatron Bridge.

NVIDIA/skills3.6k—~1.9kAutomated safety check: PassApache-2.02 days ago
226

Long-context MoE training guidance for Megatron Bridge. An agent skill from NVIDIA/skills.

NVIDIA/skills3.6k—~1.2kAutomated safety check: PassApache-2.02 days ago
227

Practical guidance for training MoE VLMs in Megatron Bridge.

NVIDIA/skills3.6k—~1.3kAutomated safety check: PassApache-2.02 days ago
228

Operational guide for enabling TP, DP, and PP communication overlap in Megatron-Bridge, including config knobs, code anchors, pitfalls, and verification.

NVIDIA/skills3.6k—~924Automated safety check: PassApache-2.02 days ago
229

Deploy ML models on Kubernetes with KServe (formerly KFServing) and NVIDIA Triton Inference Server.

BagelHole/DevOps-Security-Agent-Skills1.2k—~2.1kAutomated safety check: PassMIT4 mo ago
230

LLM and multimodal serving systems. An agent skill from uw-syfi/vibesys.

uw-syfi/vibesys105—~2.9kAutomated safety check: PassMITyesterday
231

Verl 分布式训练服务一键拉起与配置。触发场景:(1) 用户要启动 Verl 训练任务或部署 RLHF/DAPO 训练环境 (2) 在 NPU 集群上拉起 Verl 训练容器 (3) 配置 Ray 集群和 SwanLab 监控 (4) 根据 7 位二进制掩码灵活配置加速特性。支持 Qwen3-8B 等 Megatron 模型的 DAPO 训练全流程。

ascend-ai-coding/awesome-ascend-skills174—~2kAutomated safety check: PassNo licenceyesterday
232

GPU optimization for consumer NVIDIA GPUs (8-24GB VRAM) covering mixed precision, gradient checkpointing, XGBoost GPU, CuPy/cuDF migration, and torch.compile.

Mathews-Tom/armory329—~3.5kAutomated safety check: NotesMIT5 days ago
233

Use this sub-skill for Torch-TensorRT model compilation, dynamic input planning, torch.export workflows, save/load formats, raw TensorRT engines, and compile-time troubleshooting.

VectorSpaceLab/AREX-Skill331—~1.3kAutomated safety check: PassBSD-3-Clause1 mo ago
234

Use this sub-skill for Torch-TensorRT runtime performance controls, CUDA Graphs, output allocation, caches, TensorRT-RTX runtime settings, mutable modules, refit, weight streaming, and benchmark…

VectorSpaceLab/AREX-Skill331—~1kAutomated safety check: PassBSD-3-Clause1 mo ago
235

A skill your agent uses for Torch-TensorRT tasks: compiling PyTorch models with TensorRT, dynamic-shape/export workflows, runtime optimization, Triton/C++/distributed deployment, debugging…

VectorSpaceLab/AREX-Skill331—~1.5kAutomated safety check: PassBSD-3-Clause1 mo ago
236

Generate migration deliverables for bringing relevant Megatron changes into MindSpeed after branch alignment and impact mapping are complete.

ascend-ai-coding/awesome-ascend-skills174—~1.3kAutomated safety check: PassNo licenceyesterday
237

Verl 单异步 DAPO 训练配置生成器。触发场景:(1) 启动单异步 DAPO 训练 (2) 生成训练脚本 (3) 配置特性参数 (4) 训练前检查。特性策略:用户未指定时默认开启性能特性(flashattn/dynamicbatch/removepadding/gradientcheckpointing),显存特性(offload/recompute)默认关闭。OOM…

ascend-ai-coding/awesome-ascend-skills174—~1.2kAutomated safety check: PassNo licenceyesterday
238

MindSpeed-MM multimodal model suite environment setup guide for Huawei Ascend NPU.

ascend-ai-coding/awesome-ascend-skills174—~3.1kAutomated safety check: PassNo licenceyesterday
239

Universal VLM (vision-language understanding model) training guide for Huawei Ascend NPU using MindSpeed-MM.

ascend-ai-coding/awesome-ascend-skills174—~5.2kAutomated safety check: PassNo licenceyesterday
240

Generates an executable, end-to-end VERL reinforcement learning quickstart runbook for Ascend/NPU (docker image, dataset preprocessing, model setup, mainppo training, and examples/run.sh flow).

ascend-ai-coding/awesome-ascend-skills174—~592Automated safety check: PassNo licenceyesterday