Search
PyTorch · For developers
Skills
Sort:BestMost starsTrending todayTrending this weekTrending this monthNewestRecently updatedName
| # | Skill | Repository | Stars | Used in | Tokens | Auto-check | Licence | Updated |
|---|---|---|---|---|---|---|---|---|
| 97 | Train ML models on Databricks. An agent skill from databricks/databricks-agent-skills. | databricks/ | 345 | — | ~4.6k | Automated safety check: Pass | Unknown | yesterday |
| 98 | Official NVIDIA-authored guidance for navigating PhysicsNeMo — pick the model, datapipe, or example for a SciML/AI4Science task (surrogates, forecasting, downscaling, physics-informed, inverse… | NVIDIA/ | 3.6k | — | ~1.8k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 99 | Official NVIDIA-authored guidance for PhysicsNeMo ShardTensor domain parallelism — integrate domain parallelism into training/inference scripts (new or existing) with DDP or FSDP2, write and… | NVIDIA/ | 3.6k | — | ~3.5k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 100 | CLIP vision-language model for image-text retrieval, zero-shot classification, embedding extraction, ONNX export, and TensorRT deployment. | NVIDIA/ | 3.6k | — | ~4k | Automated safety check: Notes | Apache-2.0 | yesterday |
| 101 | Fine-tune any HuggingFace CV / VLM / LLM model on local NVIDIA GPUs inside an NGC PyTorch container when no dedicated TAO model skill matches. | NVIDIA/ | 3.6k | — | ~4.9k | Automated safety check: Notes | Apache-2.0 | yesterday |
| 102 | InternVideo2-CLIP L14 (TAO videoclip) for video-text retrieval, zero-shot classification, embedding extraction, LoRA fine-tuning, ONNX export, and TensorRT deployment. | NVIDIA/ | 3.6k | — | ~3.5k | Automated safety check: Notes | Apache-2.0 | yesterday |
| 103 | PyTorch-based TAO image classification. An agent skill from NVIDIA/skills. | NVIDIA/ | 3.6k | — | ~3.6k | Automated safety check: Notes | Apache-2.0 | yesterday |
| 104 | Apply existing ShapeShifter graph passes to an .onnx model via the quark-cli shapeshifter CLI or a ShapeShifter YAML. | amd/ | 182 | — | ~1.9k | Automated safety check: Pass | MIT | 13 days ago |
| 105 | Diagnose failed Quark installation, PTQ execution, script generation, or export attempts. | amd/ | 182 | — | ~1.9k | Automated safety check: Notes | MIT | 13 days ago |
| 106 | Install or verify the correct PyTorch build for a user's accelerator backend before Quark installation. | amd/ | 182 | — | ~1.6k | Automated safety check: Pass | MIT | 13 days ago |
| 107 | L3 recipe that runs a Torch LLM PTQ end-to-end for AMD Quark — for PyTorch / HuggingFace transformers models (safetensors input): quantize → validate → evaluate. | amd/ | 182 | — | ~2.6k | Automated safety check: Pass | MIT | 13 days ago |
| 108 | Generate C/C++ or CUDA code from an AI model (PyTorch, LiteRT) using MATLAB Coder or GPU Coder. | matlab/ | 1.1k | — | ~2.8k | Automated safety check: Pass | Unknown | 2 days ago |
| 109 | Creates MATLAB interfaces to Python image processing and computer vision models from GitHub repositories or pip-installable packages using MPyReq. | matlab/ | 1.1k | — | ~3.7k | Automated safety check: Pass | Unknown | 2 days ago |
| 110 | Deploy AI models to embedded hardware using MathWorks tools (MATLAB, Simulink, Embedded Coder). | matlab/ | 1.1k | — | ~4.6k | Automated safety check: Pass | Unknown | 2 days ago |
| 111 | Convert existing PyTorch Lightning training code into an NVFLARE federated job using the Lightning Client API patch, local validation, and job export; use only when the request names… | NVIDIA/ | 3.6k | — | ~3.3k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 112 | Extract Intel GPU ISA (assembly) from any XPU kernel. An agent skill from intel/torch-xpu-ops. | intel/ | 115 | — | ~2.9k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 113 | A skill your agent uses when asked to implement a fix for a root-caused failure, apply a proposed patch, or produce a staged code change from triage output. | intel/ | 115 | — | ~4.2k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 114 | A skill your agent uses when asked to verify a fix works, confirm a staged patch resolves a failure, or produce a before/after summary of a fix. | intel/ | 115 | — | ~3.3k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 115 | A skill your agent uses when asked to fix a GitHub issue end-to-end, run the agent pipeline on an issue, or process an agent:active / batch tracking issue. | intel/ | 115 | — | ~9.5k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 116 | Build PyTorch from source with Intel XPU (GPU) support. An agent skill from intel/torch-xpu-ops. | intel/ | 115 | — | ~939 | Automated safety check: Pass | Apache-2.0 | yesterday |
| 117 | Generate PyTorch release notes worksheet for XPU by searching git log between release branches. | intel/ | 115 | — | ~1.1k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 118 | 118.Quark Install Installs or verifies AMD Quark and ensures the selected Python environment has an accelerator-matched PyTorch. | amd/ | 182 | — | ~3.5k | Automated safety check: Notes | MIT | 13 days ago |
| 119 | 119.Quark Torch Ptq Runs an end-to-end AMD Quark post-training quantization workflow for PyTorch / Hugging Face LLMs: inspect a Hub or local model, choose a quantization plan, create reproducible artifacts, request… | amd/ | 182 | — | ~2.3k | Automated safety check: Pass | MIT | 13 days ago |
| 120 | Integrate a HuggingFace Computer Vision model into the NVIDIA TAO Toolkit ecosystem (tao-core config, tao-pytorch trainer, tao-deploy TensorRT pipeline). | NVIDIA/ | 3.6k | — | ~4.5k | Automated safety check: Notes | Apache-2.0 | yesterday |
| 121 | Generate and analyze T1/T2/R roofline reports for PyTorch OOB workloads comparing Intel XPU and NVIDIA CUDA. | intel/ | 115 | — | ~681 | Automated safety check: Pass | Apache-2.0 | yesterday |
| 122 | Find upstream PyTorch behavior or fixes that may require XPU parity work, validate them on XPU, and produce independently reviewed evidence. | intel/ | 115 | — | ~2k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 123 | PyTorch Geometric (PyG) for graph neural networks: node/graph classification, link prediction with GCN, GAT, GraphSAGE, GIN. | jaechang-hits/ | 374 | 1 repo | ~5.1k | Automated safety check: Pass | MIT | 12 days ago |
| 124 | A skill your agent uses when asked to reproduce a bug, verify a nightly CI failure, or confirm a failure still exists on latest source. | intel/ | 115 | — | ~7.7k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 125 | 优化实际模型推理链路,将正确性对齐、分段 profiling、显存与数据搬运、TensorRT/ONNX/PyTorch 后端、attention/kernel、FP8/compile、缓存与少步采样、质量回归、GPU 成本和服务验收串成同一实验闭环。当用户要求推理提速、降低显存或 GPU 成本、复现模型效果、定位 GPU 利用率低、优化图像/视频/扩散模型或自托管 LLM 时使用,提供… | majiayu000/ | 287 | — | ~1.1k | Automated safety check: Pass | MIT | 3 days ago |
| 126 | Techniques for reducing peak GPU memory in Megatron Bridge — expandable segments, PEFT + SP input re-gather, parallelism resizing, activation recompute, CPU offloading constraints, and common OOM… | NVIDIA/ | 3.6k | — | ~3.6k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 127 | Shallow, text-only triage of a GitHub issue on pytorch or torch-xpu-ops. | intel/ | 115 | — | ~2.5k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 128 | 128.Flash Attention Optimize transformer attention with Flash Attention — 2-4x speedup, 10-20x memory reduction for long sequences on CUDA GPUs. | AlexAI-MCP/ | 135 | — | ~1.2k | Automated safety check: Pass | MIT | 6 mo ago |
| 129 | Build a ChatGPT-like LLM from scratch using PyTorch step by step | wentorai/ | 298 | 1 repo | ~1.6k | Automated safety check: Pass | MIT | 3 mo ago |
| 130 | PyTorch Lightning framework for scalable model training and research | wentorai/ | 298 | 1 repo | ~2k | Automated safety check: Pass | MIT | 3 mo ago |
| 131 | 131.Discover ML Automatically discover machine learning and AI skills when working with machine learning, PyTorch, training, inference, RAG, embeddings, fine-tuning, LLM, DSPy, HuggingFace, or diffusion models. | rand/ | 181 | — | ~574 | Automated safety check: Pass | MIT | 7 mo ago |
| 132 | 132.GPU Optimizer GPU optimization for consumer NVIDIA GPUs (8-24GB VRAM) covering mixed precision, gradient checkpointing, XGBoost GPU, CuPy/cuDF migration, and torch.compile. | Mathews-Tom/ | 329 | — | ~3.5k | Automated safety check: Notes | MIT | 5 days ago |
| 133 | A skill your agent uses when selecting Sentence Transformers inference backends or exporting/optimizing models for PyTorch, ONNX, or OpenVINO. | VectorSpaceLab/ | 331 | — | ~766 | Automated safety check: Pass | Apache-2.0 | 1 mo ago |
| 134 | Use this sub-skill for Torch-TensorRT model compilation, dynamic input planning, torch.export workflows, save/load formats, raw TensorRT engines, and compile-time troubleshooting. | VectorSpaceLab/ | 331 | — | ~1.3k | Automated safety check: Pass | BSD-3-Clause | 1 mo ago |
| 135 | 135.Torch Tensorrt A skill your agent uses for Torch-TensorRT tasks: compiling PyTorch models with TensorRT, dynamic-shape/export workflows, runtime optimization, Triton/C++/distributed deployment, debugging… | VectorSpaceLab/ | 331 | — | ~1.5k | Automated safety check: Pass | BSD-3-Clause | 1 mo ago |
| 136 | 136.Torchvision A skill your agent uses when working with TorchVision models, weights, transforms, TVTensors, datasets, image IO, visualization utilities, vision ops, detection helpers, or official reference… | VectorSpaceLab/ | 331 | — | ~1.1k | Automated safety check: Pass | BSD-3-Clause | 1 mo ago |
| 137 | 137.Yolov5 A skill your agent uses for clone-run Ultralytics YOLOv5 workflows: detection, segmentation, classification, export, benchmarks, datasets, weights, and Flask REST serving. | VectorSpaceLab/ | 331 | — | ~1.6k | Automated safety check: Pass | AGPL-3.0 | 1 mo ago |
| 138 | Ankh 蛋白质语言模型昇腾 NPU 迁移 Skill,适用于 Ankh base/large、Ankh3 large/XL 以及同类基于 HuggingFace Transformers 与 PyTorch 的蛋白模型从 CUDA/GPU 到华为 Ascend NPU 的环境检查、代码适配、权重加载、验证脚本补齐与文档沉淀。 | ascend-ai-coding/ | 174 | — | ~2.1k | Automated safety check: Pass | No licence | yesterday |
| 139 | 昇腾 TensorFlow Community 迁移适配 Skill,适用于将基于 TensorFlow 2.x 的模型原生部署到华为 Ascend NPU,而不经过 TF 到 PyTorch 转换,覆盖 aarch64 源码编译 TF 2.6.5、tfplugin 安装、自动迁移工具使用、手动适配与精度验证。 | ascend-ai-coding/ | 174 | — | ~2.1k | Automated safety check: Pass | No licence | yesterday |
| 140 | DeepFRI 的 TensorFlow 到 PyTorch 转换与昇腾 NPU 迁移 Skill,适用于蛋白质功能预测场景下的 TF 模型分析、PyTorch 重写、权重逐层映射、NPU 推理与精度验证,尤其适合需要在 Ascend 上运行 DeepFRI CNN 或 GCN 路径时使用。 | ascend-ai-coding/ | 174 | — | ~2.4k | Automated safety check: Pass | No licence | yesterday |
| 141 | DeepFRI TensorFlow 原生昇腾 NPU 迁移 Skill,适用于不做 TF 到 PyTorch 转换、而是直接使用 TensorFlow 2.6.5 与 npudevice 在华为 Ascend 上运行 DeepFRI 的场景,覆盖源码编译、tfplugin 安装、代码适配、推理与 CPU 对比验证。 | ascend-ai-coding/ | 174 | — | ~2.1k | Automated safety check: Pass | No licence | yesterday |
| 142 | OligoFormer 昇腾 NPU 迁移 Skill,适用于将基于 PyTorch Transformer 的 siRNA 效能预测模型迁移到华为 Ascend NPU,覆盖环境搭建、RNA-FM 依赖安装、代码适配、推理验证与可选训练流程。 | ascend-ai-coding/ | 174 | — | ~1.3k | Automated safety check: Pass | No licence | yesterday |
| 143 | ProteinBERT 昇腾 NPU 部署与迁移 Skill,适用于将 TensorFlow 或 Keras 版 ProteinBERT 转成基于 PyTorch 与 torchnpu 的实现,覆盖权重转换、embedding 提取、微调训练、注意力可视化和 GPU 与 NPU 精度验证。 | ascend-ai-coding/ | 174 | — | ~1.9k | Automated safety check: Pass | No licence | yesterday |
| 144 | TensorFlow 或 Keras 模型改写到 PyTorch 的通用 Skill,适用于在华为 Ascend NPU 或其他依赖 PyTorch 生态的平台上完成层级映射、权重转换、逐层数值验证和端到端精度对比,尤其适合 ProteinBERT、DeepFRI 这类科学模型的跨框架迁移。 | ascend-ai-coding/ | 174 | — | ~3.4k | Automated safety check: Pass | No licence | yesterday |