Search

CUDA

282 skills found, page 5.
Search results
#SkillRepositoryStarsUsed inTokensAuto-checkLicenceUpdated
193

Install or repair the FoundationPose perception pipeline and build its FoundationStereo TensorRT engines.

NVIDIA/skills3.6k—~1.5kAutomated safety check: PassApache-2.0yesterday
194
194.Nv Generate MrOfficial

Used for generating synthetic body MRI volumes with NV-Generate-CTMR rflow-mr.

NVIDIA/skills3.6k—~2kAutomated safety check: NotesApache-2.0yesterday
195

A skill your agent uses for NVIDIA-related requests where an NVIDIA skill might help, even if the user did not ask for a skill.

NVIDIA/skills3.6k—~2.1kAutomated safety check: PassApache-2.0yesterday
196

NV-Tesseract AD Diffusion — diffusion-based anomaly detection and fine-tuning for multivariate time series.

NVIDIA/skills3.6k—~2.9kAutomated safety check: NotesApache-2.0yesterday
197

NV-Tesseract Forecasting — transformer-based multivariate time series forecasting with DARR (context-enhanced kNN retrieval), interpretability, and fine-tuning.

NVIDIA/skills3.6k—~3.2kAutomated safety check: NotesApache-2.0yesterday
198

cuOpt REST server — start server, endpoints, Python/curl client examples.

NVIDIA/skills3.6k—~1.5kAutomated safety check: PassApache-2.0yesterday
199

Maps every HOT-Step CPP feature to its route file, service, UI folder, and engine subsystem, including port topology and the browser-to-engine request path.

scragnog/HOT-Step-CPP174—~5.4kAutomated safety check: NotesMIT2 days ago
200

Build a Quark ONNX PTQ quantization plan from modelanalysis.json and user intent.

amd/Quark182—~4.8kAutomated safety check: PassMIT13 days ago
201

Diagnose failed Quark installation, PTQ execution, script generation, or export attempts.

amd/Quark182—~1.9kAutomated safety check: NotesMIT13 days ago
202

Install or verify the correct PyTorch build for a user's accelerator backend before Quark installation.

amd/Quark182—~1.6kAutomated safety check: PassMIT13 days ago
203

Diagnoses and resolves MCP server registration failures, GPU detection, BigQuery authentication, index build failures, import errors, search quality issues, and performance problems.

RobThePCGuy/Claude-Patent-Creator196—~1.2kAutomated safety check: PassMIT4 days ago
204

Diagnose CUDA "illegal instruction" / kernel crashes on Triton kernels that reference to TMA loads or stores (maketensordescriptor, TensorDescriptor, descriptor.load, descriptor.store…

facebookexperimental/triton201—~1.1kAutomated safety check: PassMITyesterday
205

Generate C/C++ or CUDA code from an AI model (PyTorch, LiteRT) using MATLAB Coder or GPU Coder.

matlab/matlab-agentic-toolkit1.1k—~2.8kAutomated safety check: PassUnknown2 days ago
206

Generate, verify, refine, and accelerate C/C++ or CUDA code from MATLAB with MATLAB Coder, Embedded Coder, GPU Coder, or MATLAB Test.

matlab/matlab-agentic-toolkit1.1k—~4.2kAutomated safety check: PassUnknown2 days ago
207

Deploy AI models to embedded hardware using MathWorks tools (MATLAB, Simulink, Embedded Coder).

matlab/matlab-agentic-toolkit1.1k—~4.6kAutomated safety check: PassUnknown2 days ago
208

A skill your agent uses when the operator is authoring, building, loading, or debugging a custom doca-bench plug-in — a versioned shared library with DOCAEXPERIMENTAL-marked C entry points that…

NVIDIA/skills3.6k—~4kAutomated safety check: PassApache-2.0yesterday
209
209.Doca GpiOfficial

A skill your agent uses for hands-on DOCA GPI programming — wiring a GPU-Packet-Initiator context so a CUDA kernel drives RDMA queues directly from GPU memory without host CPU mediation.

NVIDIA/skills3.6k—~3.9kAutomated safety check: PassApache-2.0yesterday
210
210.Doca GpunetioOfficial

A skill your agent uses when the user is doing hands-on DOCA GPUNetIO programming — wiring a CUDA kernel on an NVIDIA GPU to a doca-eth queue via docagpuethrxq / docagpuethtxq, standing up the…

NVIDIA/skills3.6k—~3.7kAutomated safety check: PassApache-2.0yesterday
211

A skill your agent uses when the user is building, running, or interpreting the doca/tools/gpunetioibwritebw client+server benchmark — a CUDA kernel on the client posts RDMA WRITE work requests…

NVIDIA/skills3.6k—~4.2kAutomated safety check: PassApache-2.0yesterday
212

A skill your agent uses when the user is measuring GPU-kernel-initiated RDMA WRITE latency through doca-gpunetio — building and running the gpunetioibwritelat client + server pair under…

NVIDIA/skills3.6k—~3.8kAutomated safety check: PassApache-2.0yesterday
213

Guide installing Earth2Studio via uv or pip, selecting model extras, and configuring the environment.

NVIDIA/skills3.6k—~1.6kAutomated safety check: NotesApache-2.0yesterday
214

Install Holoscan SDK v4.3+ via Conda in a CUDA 13 environment.

NVIDIA/skills3.6k—~2kAutomated safety check: PassApache-2.0yesterday
215

Install Holoscan SDK via the NGC Docker container. An agent skill from NVIDIA/skills.

NVIDIA/skills3.6k—~1.9kAutomated safety check: PassApache-2.0yesterday
216

Install Holoscan SDK natively on Ubuntu via apt. An agent skill from NVIDIA/skills.

NVIDIA/skills3.6k—~1.6kAutomated safety check: NotesApache-2.0yesterday
217

Install Holoscan SDK Python wheel via pip into a venv. An agent skill from NVIDIA/skills.

NVIDIA/skills3.6k—~1.6kAutomated safety check: NotesApache-2.0yesterday
218
218.Holoscan SetupOfficial

Guides Holoscan SDK installation: inspects the host, assesses platform compatibility, recommends an install method, and delegates to the matching install skill.

NVIDIA/skills3.6k—~2.5kAutomated safety check: PassApache-2.0yesterday
219

Validate and use selective and full activation recompute in Megatron Bridge to reduce GPU memory usage at the cost of extra compute.

NVIDIA/skills3.6k—~4.7kAutomated safety check: PassApache-2.0yesterday
220
220.Nv Reason CxrOfficial

Used for command-shape or live NV-Reason-CXR chest X-ray reasoning smoke tests.

NVIDIA/skills3.6k—~3.9kAutomated safety check: NotesApache-2.0yesterday
221
221.Nv Segment CtOfficial

Used for running NV-Segment-CT VISTA3D on CT NIfTI volumes and recording label-map evidence.

NVIDIA/skills3.6k—~2.1kAutomated safety check: NotesApache-2.0yesterday
222
222.Nv Segment CtmrOfficial

Used for running NV-Segment-CTMR on CT or MRI NIfTI volumes and recording label-map evidence.

NVIDIA/skills3.6k—~2.3kAutomated safety check: NotesApache-2.0yesterday
223
223.Paidf AugmentationOfficial

A skill your agent uses when authoring or validating PAIDF augmentation YAML configs, or running remote Cosmos Transfer (including Cosmos3 WSM controls), Cosmos Predict, image-edit, or…

NVIDIA/skills3.6k—~5.3kAutomated safety check: PassApache-2.0yesterday
224

How to create a pull request for the intel/torch-xpu-ops repository.

intel/torch-xpu-ops115—~1.2kAutomated safety check: PassApache-2.0yesterday
225

Official NVIDIA-authored guidance for NVIDIA cuDF GPU DataFrames, pandas acceleration, dask-cuDF, ETL, joins, groupby, CSV/Parquet I/O, nullable semantics, and multi-GPU DataFrame workloads.

NVIDIA/skills3.6k—~2.6kAutomated safety check: PassApache-2.0yesterday
226
226.Cudaq ImportingOfficial

A skill your agent uses when porting circuits from another framework (e.g.

NVIDIA/skills3.6k—~1.9kAutomated safety check: PassApache-2.0yesterday
227
227.Cuopt InstallOfficial

Install cuOpt for Python, C, or server via pip, conda, or Docker; verify the install.

NVIDIA/skills3.6k—~1.1kAutomated safety check: PassApache-2.0yesterday
228

Create, refine, or fix NVIDIA voice agents (Cascaded or Omni) with Pipecat or LiveKit.

NVIDIA/skills3.6k—~844Automated safety check: PassApache-2.0yesterday
229
229.RAG BlueprintOfficial

NVIDIA RAG Blueprint — deploy, configure, troubleshoot, and manage.

NVIDIA/skills3.6k—~2.8kAutomated safety check: NotesApache-2.0yesterday
230

How to swap the DeepStream CV detection model in the VSS Alerts Blueprint verification (2dcv) mode - covers ONNX export, custom bbox parsers, compose mount gotchas, nvinfer config, runtime TRT…

NVIDIA/skills3.6k—~4.5kAutomated safety check: NotesApache-2.0yesterday
231
231.Oob Perf AnalysisOfficial

Generate and analyze T1/T2/R roofline reports for PyTorch OOB workloads comparing Intel XPU and NVIDIA CUDA.

intel/torch-xpu-ops115—~681Automated safety check: PassApache-2.0yesterday
232

Build and run LAMMPS molecular dynamics with isolated MLIP-specific binaries (MACE, MatGL/CHGNet, FairChem) to avoid Python and Torch stack conflicts.

learningmatter-mit/AtomisticSkills176—~1.2kAutomated safety check: PassMIT3 days ago
233

Generate inorganic material structures using MatterGen, a diffusion-based generative model.

learningmatter-mit/AtomisticSkills176—~1.8kAutomated safety check: PassMIT3 days ago
234

Run and diagnose governed ASE plus fairchem UMA molecular dynamics on standardized adsorbed slabs, including NVT/NVE surface trajectories, pre-relaxation, periodic-boundary and collision checks…

Tai609/NebulaMat100—~1.6kAutomated safety check: PassUnknown1 mo ago
235

How HOT-Step's custom flash-attention training ops (GGMLOPFLASHATTNTRAIN/BACK) work, what the AS1.5 DiT trainer campaign proved and disproved, and the exact contract for porting flash mode to the…

scragnog/HOT-Step-CPP174—~5.4kAutomated safety check: PassMIT2 days ago
236

Use a local QMD knowledge base through UXC over MCP stdio, with daemon-backed session reuse and typed retrieval flows that avoid repeated model warmup and unnecessary query-expansion latency.

holon-run/uxc116—~1.3kAutomated safety check: PassMIT26 days ago
237

Build Holoscan SDK from source via the in-tree ./run script.

NVIDIA/skills3.6k—~1.5kAutomated safety check: NotesApache-2.0yesterday
238

Validate and use CPU offloading in Megatron Bridge, including layer-level activation offloading and fractional optimizer state offloading with HybridDeviceOptimizer.

NVIDIA/skills3.6k—~2.3kAutomated safety check: PassApache-2.0yesterday
239

Validate and use CUDA graph capture in Megatron Bridge, including local full-iteration graphs and Transformer Engine scoped graphs for attention, MLP, and MoE modules.

NVIDIA/skills3.6k—~3.5kAutomated safety check: PassApache-2.0yesterday
240

Validate and use MoE expert-parallel communication overlap in Megatron-Bridge, including overlapmoeexpertparallelcomm, delaywgradcompute, and flex dispatcher backends such as DeepEP and HybridEP.

NVIDIA/skills3.6k—~3.5kAutomated safety check: PassApache-2.0yesterday