Search

DevOps & Cloud · NVIDIA AI Platform · By NVIDIA

55 skills found.
Search results
#SkillRepositoryStarsUsed inTokensAuto-checkLicenceUpdated
1

Walks an agent through working inside the Megatron-LM CI container and changing dependencies with uv, so lock files resolve the same locally and in CI.

NVIDIA/Megatron-LM18k—~2.6kAutomated safety check: PassApache-2.0today
2

Moves Megatron-LM CI to a newer NVIDIA PyTorch base image, updating both the GitHub and GitLab pins together and handling the CI follow-up.

NVIDIA/Megatron-LM18k—~2.8kAutomated safety check: PassApache-2.0today
3

Investigates a failing GitHub Actions run or job for Megatron-LM, finds the root cause plus the PR and test author involved, and files a structured bug issue.

NVIDIA/Megatron-LM18k—~1.6kAutomated safety check: PassApache-2.0today
4

Install NVIDIA NIM Operator on Kubernetes with prerequisite checks, optional NVIDIA GPU Operator dependency installation, public or local Helm chart selection, optional Dynamo support, and optional…

NVIDIA/k8s-nim-operator159—~4.7kAutomated safety check: PassApache-2.05 days ago
5

Safely uninstall NVIDIA NIM Operator from Kubernetes with inventory checks, explicit approval gates for destructive actions, optional custom resource cleanup, optional CRD removal, and…

NVIDIA/k8s-nim-operator159—~3.6kAutomated safety check: PassApache-2.05 days ago
6

A skill your agent uses when validating DCGM Exporter in a local GPU-backed k3d/Kubernetes environment.

NVIDIA/dcgm-exporter1.9k—~116Automated safety check: PassApache-2.021 days ago
7

Multi-agent PR review using Claude Code, Codex, and CodeRabbit.

NVIDIA/aicr440—~15kAutomated safety check: PassApache-2.0today
8

Shows how to launch distributed Megatron-LM training on a SLURM cluster: sbatch skeleton, torch.distributed.run setup, CUDA_DEVICE_MAX_CONNECTIONS rules and failure diagnosis.

NVIDIA/Megatron-LM18k—~1.8kAutomated safety check: PassApache-2.0today
9

Use BEFORE running a full CompileIQ search. An agent skill from NVIDIA/CompileIQ.

NVIDIA/CompileIQ138—~2.1kAutomated safety check: NotesApache-2.017 days ago
10

Start up, tear down, and configure the local Kubernetes development environment for OpenShell.

NVIDIA/OpenShell16k—~4.9kAutomated safety check: PassApache-2.0today
11

A skill your agent uses when analyzing an AICR snapshot YAML file, reviewing cluster state, comparing provider characteristics, extracting GPU/network topology insights, or generating a cluster…

NVIDIA/aicr440—~3.5kAutomated safety check: PassApache-2.0today
12

A skill your agent uses when migrating applications, examples, integrations, documentation, manifests, or repository code from NeMo Flow to NeMo Relay across Python, Rust, Node.js, Go, C FFI, CLI…

NVIDIA/NeMo-Relay192—~1.8kAutomated safety check: PassApache-2.0yesterday
13

A skill your agent uses when reviewing the weekly AICR component drift report — the Slack digest and drift-report.json artifact produced by Registry Drift Report (registry-drift.yaml) listing which…

NVIDIA/aicr440—~2.8kAutomated safety check: PassApache-2.0today
14

Classify one failed NemoClaw GitHub Actions job using bounded, redacted logs and optional retained artifacts.

NVIDIA/NemoClaw23k—~806Automated safety check: PassApache-2.0today
15
15.Evo2 NimOfficial

Generate and analyze DNA sequences using NVIDIA's Evo 2 BioNeMo NIM microservice.

NVIDIA/skills3.6k1 repo~2.4kAutomated safety check: NotesApache-2.0yesterday
16
16.Msa Search NimOfficial

Generate multiple sequence alignments (MSAs) for protein sequences using the ColabFold MSA-Search NIM.

NVIDIA/skills3.6k1 repo~4.6kAutomated safety check: NotesApache-2.0yesterday
17
17.Proteinmpnn NimOfficial

Run ProteinMPNN inverse folding via NVIDIA NIM to design protein sequences for a target backbone.

NVIDIA/skills3.6k1 repo~2kAutomated safety check: NotesApache-2.0yesterday
18

Verify whether a tagged NeMo Relay release reached GitHub Actions, tagged Go module source, crates.io, PyPI, and npm.

NVIDIA/NeMo-Relay192—~415Automated safety check: PassApache-2.0yesterday
19
19.Diffdock NimOfficial

Run DiffDock molecular docking via NVIDIA NIM to predict small-molecule binding poses against protein targets.

NVIDIA/skills3.6k1 repo~1.1kAutomated safety check: NotesApache-2.0yesterday
20

Validate that a Dynamo deployment's NIXL/UCX/NCCL interconnect is ready for disaggregated serving over RDMA/NVLink.

NVIDIA/skills3.6k1 repo~1.6kAutomated safety check: PassApache-2.0yesterday
21
21.Genmol NimOfficial

Generate novel drug-like molecules using the GenMol NIM microservice.

NVIDIA/skills3.6k1 repo~1.4kAutomated safety check: NotesApache-2.0yesterday
22
22.Molmim NimOfficial

A skill your agent uses for MolMIM, NVIDIA's BioNeMo NIM microservice for small-molecule latent-space generation and optimization.

NVIDIA/skills3.6k1 repo~1.9kAutomated safety check: NotesApache-2.0yesterday
23
23.Openfold2 NimOfficial

A skill your agent uses for OpenFold2, NVIDIA's BioNeMo NIM microservice for monomer protein structure prediction.

NVIDIA/skills3.6k1 repo~1.8kAutomated safety check: NotesApache-2.0yesterday
24
24.Openfold3 NimOfficial

A skill your agent uses for OpenFold3, NVIDIA's BioNeMo NIM microservice for biomolecular structure prediction.

NVIDIA/skills3.6k1 repo~1.9kAutomated safety check: NotesApache-2.0yesterday
25
25.Rfdiffusion NimOfficial

Run RFDiffusion protein backbone design via NVIDIA NIM. An agent skill from NVIDIA/skills.

NVIDIA/skills3.6k1 repo~1.3kAutomated safety check: NotesApache-2.0yesterday
26
26.Kermt SetupOfficial

Bootstrap the KERMT agent environment — verify host docker + nvidia-container-toolkit, build the kermt:latest image from the repo's Dockerfile if it doesn't yet exist, and run a GPU smoke test…

NVIDIA/skills3.6k1 repo~1.7kAutomated safety check: PassApache-2.0yesterday
27

Select, validate, patch, and deploy existing NVIDIA Dynamo Kubernetes recipes.

NVIDIA/skills3.6k—~1.8kAutomated safety check: PassApache-2.0yesterday
28
28.Tao Run AutomlOfficial

Run container-backed AutoML / hyperparameter optimization (HPO) for NVIDIA TAO networks using AutoMLRunner.

NVIDIA/skills3.6k—~5kAutomated safety check: NotesApache-2.0yesterday
29

Host setup for TAO GPU backends. An agent skill from NVIDIA/skills.

NVIDIA/skills3.6k—~3.4kAutomated safety check: NotesApache-2.0yesterday
30

A skill your agent uses for VSS alert workflows — real-time monitoring, Alert-Bridge subscriptions, Slack notifications, incident queries, camera onboarding.

NVIDIA/skills3.6k—~4.5kAutomated safety check: NotesApache-2.0yesterday
31
31.Boltz2 NimOfficial

Use Boltz2 NIM for biomolecular structure prediction and binding affinity.

NVIDIA/skills3.6k1 repo~1.4kAutomated safety check: NotesApache-2.0yesterday
32

Model and publish semantic definitions in Auto Ontology. An agent skill from NVIDIA/skills.

NVIDIA/skills3.6k—~1.6kAutomated safety check: PassApache-2.0yesterday
33

Set up or troubleshoot the Auto Ontology runtime. An agent skill from NVIDIA/skills.

NVIDIA/skills3.6k—~2.7kAutomated safety check: NotesApache-2.0yesterday
34

A skill your agent uses to deploy and operate a CollectX (clx) based DOCA telemetry collector on a host or BlueField — wiring providers / counters into the collector, running the collection daemon…

NVIDIA/skills3.6k—~2.8kAutomated safety check: PassApache-2.0yesterday
35

A skill your agent uses when the user is hands-on deploying an in-bundle DOCA service container (Argus, DMS, Firefly, or UROM service) on a BlueField — kubelet standalone watching a static-pod…

NVIDIA/skills3.6k—~2.5kAutomated safety check: PassApache-2.0yesterday
36

Performs gap analysis on NVIDIA TAO VCN Classify (Visual Component Net) experiments by invoking the pinned TAO data-services container directly via docker run … gapanalysis vcnaoi … — picks the…

NVIDIA/skills3.6k—~4.3kAutomated safety check: NotesApache-2.0yesterday
37

The mandatory pre-launch gate and four-verb execution contract for every TAO workflow or action.

NVIDIA/skills3.6k—~4.5kAutomated safety check: NotesApache-2.0yesterday
38

Pose classification using ST-GCN (Spatial Temporal Graph Convolutional Network).

NVIDIA/skills3.6k—~3.8kAutomated safety check: NotesApache-2.0yesterday
39

A skill your agent uses to run top-level VSS fusion search on archived video, or to ingest video files / RTSP streams for search.

NVIDIA/skills3.6k—~4.2kAutomated safety check: PassApache-2.0yesterday
40

cuOpt REST server — start server, endpoints, Python/curl client examples.

NVIDIA/skills3.6k—~1.5kAutomated safety check: PassApache-2.0yesterday
41

Launch AutoMagicCalib microservice and web UI from NGC release images via Docker Compose.

NVIDIA/skills3.6k—~3.9kAutomated safety check: NotesApache-2.0yesterday
42
42.Doca ArgusOfficial

A skill your agent uses when the user is deploying or operating the DOCA Argus Service — the packaged BlueField-side runtime-security container that watches the BlueField and attached host for…

NVIDIA/skills3.6k—~4.8kAutomated safety check: PassApache-2.0yesterday
43
43.Doca DmsOfficial

Operate NVIDIA DOCA Management Service (dmsd + dmspe) on a BlueField, Arm/x86 host, or Kubernetes pod: choose deployment and authentication, configure -allowedusers and dmsgroup, use gNMI…

NVIDIA/skills3.6k—~2.8kAutomated safety check: PassApache-2.0yesterday
44

Install Holoscan SDK via the NGC Docker container. An agent skill from NVIDIA/skills.

NVIDIA/skills3.6k—~1.9kAutomated safety check: PassApache-2.0yesterday
45

A skill your agent uses when integrating NVIDIA NeMo Fabric into a consumer application, service, evaluation harness, or platform through the typed Python SDK — translating the consumer's own…

NVIDIA/skills3.6k—~5.8kAutomated safety check: PassApache-2.0yesterday
46

A skill your agent uses when reading video-analytics metrics, incidents, alerts, and sensor data via the VA-MCP server (port 9901).

NVIDIA/skills3.6k—~2.3kAutomated safety check: PassApache-2.0yesterday
47
47.Cuopt InstallOfficial

Install cuOpt for Python, C, or server via pip, conda, or Docker; verify the install.

NVIDIA/skills3.6k—~1.1kAutomated safety check: PassApache-2.0yesterday
48

How to swap the DeepStream CV detection model in the VSS Alerts Blueprint verification (2dcv) mode - covers ONNX export, custom bbox parsers, compose mount gotchas, nvinfer config, runtime TRT…

NVIDIA/skills3.6k—~4.5kAutomated safety check: NotesApache-2.0yesterday