Topic · AI & LLM Engineering

Best GPU and accelerator computing skills, page 4

Skills #145–176 of 176, ranked by score.

GPU and accelerator computing skills, ranked

Ranked by score. Sort bymost stars,trending,newest,recently updated

GPU and accelerator computing skills, ranked
#SkillRepositoryStarsUsed inTokensAuto-checkLicenceUpdated
145

A skill your agent uses when you need to print Jetson BSP info (L4T version, board configs, rootfs state) from a LinuxforTegra root on the host PC.

NVIDIA/skills3.5k—~1.3kAutomated safety check: PassApache-2.0today
146

On-demand GPU cloud instances for ML training. An agent skill from Luciole-Studio/Misaka-Agent.

Luciole-Studio/Misaka-Agent1392 repos~3kAutomated safety check: WarnMITtoday
147

Enable MIPI/GMSL camera sensors on a Jetson Thor or Orin custom carrier by rendering a kernel-DT overlay from the in-tree sensor DTSI.

NVIDIA/skills3.5k—~2.4kAutomated safety check: PassApache-2.0today
148

Per-controller PCIe enable / disable / lanes / link-speed for a Jetson Thor or Orin custom carrier via ODMDATA + kernel-DT overlay.

NVIDIA/skills3.5k—~2.3kAutomated safety check: PassApache-2.0today
149

Configure Jetson UPHY lane allocation (uphy0/uphy1-config) on Orin/Thor custom carriers.

NVIDIA/skills3.5k—~2.4kAutomated safety check: PassApache-2.0today
150

Enable/disable Jetson USB2/USB3 SS ports via kernel-DT overlay.

NVIDIA/skills3.5k—~2.1kAutomated safety check: PassApache-2.0today
151
151.Jetson Init SourceOfficial

Set up the BSP source workspace: LinuxforTegra overlay tracker, bspsources, Crosstool-NG toolchain.

NVIDIA/skills3.5k—~4.3kAutomated safety check: PassApache-2.0today
152
152.Jetson Init TargetOfficial

Author a new Jetson target-platform profile (referencedevkit + optional customcarrier) and update the active pointer.

NVIDIA/skills3.5k—~4.8kAutomated safety check: PassApache-2.0today
153

Convert single-node scripts to multi-node Slurm sbatch jobs and debug common multi-node failures.

NVIDIA/skills3.5k—~3.9kAutomated safety check: PassApache-2.0today
154
154.Runpod

A skill your agent uses when running GPU compute on RunPod and deciding between Pods (hourly, always-on) and Serverless (per-second, autoscaling) for training, fine-tuning or inference — serverless…

ericrisco/rsc-harness167—~2.8kAutomated safety check: PassMITtoday
155

Compare Triton, TLX, and inductor kernel numerics across A/B compiler builds with per-config isolation tests and grouped impact reporting.

facebookexperimental/triton201—~792Automated safety check: PassMITtoday
156
156.Modal

Cloud computing platform for running Python on GPUs and serverless infrastructure.

BioTender-max/awesome-bio-agent-skills197—~3.1kAutomated safety check: NotesApache-2.03 mo ago
157

Bootstrap a custom carrier board by forking carrier files and scaffolding a DT overlay from the reference devkit.

NVIDIA/skills3.5k—~4.2kAutomated safety check: PassApache-2.0today
158
158.Jetson Generate KbOfficial

Build a per-target knowledge-base markdown next to the active profile by walking the BSP root and source tree.

NVIDIA/skills3.5k—~3.8kAutomated safety check: PassApache-2.0today
159
159.Jetson Quick StartOfficial

Entry skill for Jetson / IGX BSP customization. An agent skill from NVIDIA/skills.

NVIDIA/skills3.5k—~4.7kAutomated safety check: PassApache-2.0today
160
160.Jetson Set TargetOfficial

Switch the active Jetson target-platform pointer to an existing profile YAML.

NVIDIA/skills3.5k—~1.7kAutomated safety check: PassApache-2.0today
161

Techniques for reducing peak GPU memory in Megatron Bridge — expandable segments, PEFT + SP input re-gather, parallelism resizing, activation recompute, CPU offloading constraints, and common OOM…

NVIDIA/skills3.5k—~3.6kAutomated safety check: PassApache-2.0today
162

Diagnose and fix CoreWeave GPU scheduling, pod, and networking errors.

jeremylongshore/tons-of-skills-marketplace2.8k—~1.1kAutomated safety check: PassMITtoday
163

Run distributed GPU training jobs on CoreWeave with multi-node PyTorch.

jeremylongshore/tons-of-skills-marketplace2.8k—~1.2kAutomated safety check: PassMITtoday
164

Optimize CoreWeave GPU cloud costs with right-sizing and scheduling.

jeremylongshore/tons-of-skills-marketplace2.8k—~1.3kAutomated safety check: PassMITtoday
165

GPU optimization for consumer NVIDIA GPUs (8-24GB VRAM) covering mixed precision, gradient checkpointing, XGBoost GPU, CuPy/cuDF migration, and torch.compile.

Mathews-Tom/armory328—~3.5kAutomated safety check: NotesMIT2 days ago
166

GPU-accelerate Python code using CuPy, Numba CUDA, Warp, cuDF, cuML, cuGraph, KvikIO, cuCIM, cuxfilter, cuVS, cuSpatial, and RAFT.

majiayu000/claude-skill-registry6661 repo~8.5kAutomated safety check: PassMITtoday
167
167.Cuda

CUDA C/C++ skill for NVIDIA GPU kernel programming. An agent skill from mohitmishra786/low-level-dev-skills.

mohitmishra786/low-level-dev-skills253—~1.9kAutomated safety check: PassMIT3 mo ago
168

CUDA debugging skill for GPU program correctness. An agent skill from mohitmishra786/low-level-dev-skills.

mohitmishra786/low-level-dev-skills253—~1.5kAutomated safety check: PassMIT3 mo ago
169

CUDA profiling skill for NVIDIA GPU performance analysis. An agent skill from mohitmishra786/low-level-dev-skills.

mohitmishra786/low-level-dev-skills253—~1.6kAutomated safety check: NotesMIT3 mo ago
170

HIP and ROCm skill for AMD GPU programming. An agent skill from mohitmishra786/low-level-dev-skills.

mohitmishra786/low-level-dev-skills253—~1.6kAutomated safety check: NotesMIT3 mo ago
171

Triton language skill for Python GPU kernel authoring. An agent skill from mohitmishra786/low-level-dev-skills.

mohitmishra786/low-level-dev-skills253—~1.8kAutomated safety check: PassMIT3 mo ago
172

Expert integration with NVIDIA GPU-accelerated math libraries.

majiayu000/claude-skill-registry6661 repo~2.4kAutomated safety check: NotesMITtoday
173

High-performance kernel template libraries and DSLs. An agent skill from majiayu000/claude-skill-registry.

majiayu000/claude-skill-registry6661 repo~2.6kAutomated safety check: NotesMITtoday
174

NVIDIA Collective Communications Library integration for multi-GPU operations.

majiayu000/claude-skill-registry6661 repo~1.9kAutomated safety check: NotesMITtoday
175
175.Jetson Link DocsOfficial

Bind pre-downloaded Jetson reference docs (developer guide, design guide, pinmux, schematics) into the active profile documents block.

NVIDIA/skills3.5k—~3.3kAutomated safety check: PassApache-2.0today
176

Audit, prepare, and deploy PAIDF Orchestration on a Kubernetes GPU cluster - single-GPU H100/L40S hosts, managed Kubernetes, kubeadm, and similar.

NVIDIA/skills3.5k—~3.8kAutomated safety check: WarnApache-2.0today