Search
C++ · Deep learning
Skills
Sort:BestMost starsTrending todayTrending this weekTrending this monthNewestRecently updatedName
| # | Skill | Repository | Stars | Used in | Tokens | Auto-check | Licence | Updated |
|---|---|---|---|---|---|---|---|---|
| 1 | A skill your agent uses when needing to compile, rebuild, or install Paddle from source after code changes. | PaddlePaddle/ | 24k | — | ~1k | Automated safety check: Pass | Apache-2.0 | 8 days ago |
| 2 | A skill your agent uses when navigating Paddle eager-mode (dynamic graph) source code, tracing forward/backward execution, debugging autograd issues, understanding PyLayer, or investigating… | PaddlePaddle/ | 24k | — | ~562 | Automated safety check: Pass | Apache-2.0 | 8 days ago |
| 3 | 3.Onnxtxt Read or write ONNX text format ("onnxtxt"). An agent skill from onnx/onnx. | onnx/ | 22k | — | ~1.3k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 4 | Builds ExecuTorch from source: the Python package, C++ runtime, model runners, Android and iOS cross-compilation and backend-specific builds, with environment checks. | pytorch/ | 5.1k | — | ~2.3k | Automated safety check: Notes | Unknown | today |
| 5 | 5.Ako4all Drive an agentic loop that iteratively optimizes a GPU kernel for maximum speedup. | TongmingLAIC/ | 369 | — | ~4k | Automated safety check: Pass | MIT | 23 days ago |
| 6 | A skill your agent uses when writing, debugging, porting, reviewing, or optimizing CUDA C++ or PTX kernels; investigating CUDA Runtime or Driver API behavior; profiling kernels with Nsight Systems… | vipshop/ | 1.3k | — | ~2.3k | Automated safety check: Pass | Apache-2.0 | 9 days ago |
| 7 | PaddlePaddle (飞桨) C++ 算子开发指南。提供从 YAML 配置、InferMeta 函数、Kernel 实现、Python API 封装、单元测试到编译验证的完整算子开发流程指导。在以下场景使用此 skill:(1) 为 Paddle 框架新增 C++ 算子 (2) 修改或调试已有 Paddle 算子 (3) 编写算子的 YAML… | PaddlePaddle/ | 24k | — | ~1.3k | Automated safety check: Pass | Apache-2.0 | 8 days ago |
| 8 | Deploy AI models to embedded hardware using MathWorks tools (MATLAB, Simulink, Embedded Coder). | matlab/ | 181 | 1 repo | ~3.4k | Automated safety check: Pass | Unknown | 27 days ago |
| 9 | Helps build, test and extend the Qualcomm AI Engine Direct (QNN) backend in ExecuTorch, with routes for new ops, model export, Buck-vs-CMake parity fixes and per-layer accuracy debugging. | pytorch/ | 5.1k | — | ~1.8k | Automated safety check: Pass | Unknown | today |
| 10 | 将原生 PyTorch 自定义算子库、Torch extension、生态库(TorchCodec/FlashInfer/DeepEP 等)以及 Kernel DSL 生态(Triton/TileLang/TVM FFI 等)以最小修改方式接入 PaddlePaddle。遇到以下场景务必使用:迁移外部算子库到 Paddle;分析 PFCCLab fork 与上游的兼容差异;处理… | PaddlePaddle/ | 24k | — | ~883 | Automated safety check: Pass | Apache-2.0 | 8 days ago |
| 11 | Convert PyTorch ATDISPATCH macros to ATDISPATCHV2 format in ATen C++ code. | intel/ | 115 | 3 repos | ~2.2k | Automated safety check: Pass | Apache-2.0 | today |
| 12 | 12.AI Review 使用 PaddlePaddle 仓库规则评审 Pull Request 和全仓库代码变更,覆盖正确性、兼容性、算子、分布式、数值、性能、安全、测试、构建和 PR 信息。当需要审查 Paddle 的代码、测试、算子 YAML、C++/CUDA/XPU kernel、Python API、分布式逻辑或 CI 配置时使用。 | PaddlePaddle/ | 24k | — | ~303 | Automated safety check: Pass | Apache-2.0 | 8 days ago |
| 13 | Install or verify the AMD Quark package and its dependencies. | amd/ | 181 | — | ~1.8k | Automated safety check: Notes | MIT | 10 days ago |
| 14 | 14.Llama Nnc Develop and profile the edge-e3 PyTorch-to-NNC flow, including nnc/compiler.py lowering and generated ABI, example/llama smoke models, cpp/libnn runtime operators, BF16 correctness checks, and… | exeex/ | 110 | — | ~1.1k | Automated safety check: Pass | Apache-2.0 | 12 days ago |
| 15 | Run, extend, debug, or review the public edge-e3 bare-metal software harness, including encrypted Verilator builds, hello and tensor examples, all example/llama/model smoke cases, PyTorch BF16… | exeex/ | 110 | — | ~1k | Automated safety check: Pass | Apache-2.0 | 12 days ago |
| 16 | Generate C/C++ or CUDA code from an AI model (PyTorch, LiteRT) using MATLAB Coder or GPU Coder. | matlab/ | 1.1k | — | ~2.8k | Automated safety check: Pass | Unknown | yesterday |
| 17 | Extract Intel GPU ISA (assembly) from any XPU kernel. An agent skill from intel/torch-xpu-ops. | intel/ | 115 | — | ~2.9k | Automated safety check: Pass | Apache-2.0 | today |
| 18 | Build PyTorch from source with Intel XPU (GPU) support. An agent skill from intel/torch-xpu-ops. | intel/ | 115 | — | ~939 | Automated safety check: Pass | Apache-2.0 | today |
| 19 | Deploy AI models to embedded hardware using MathWorks tools (MATLAB, Simulink, Embedded Coder). | majiayu000/ | 666 | 1 repo | ~4.6k | Automated safety check: Pass | MIT | today |
| 20 | A skill your agent uses for Torch-TensorRT tasks: compiling PyTorch models with TensorRT, dynamic-shape/export workflows, runtime optimization, Triton/C++/distributed deployment, debugging… | VectorSpaceLab/ | 328 | — | ~1.5k | Automated safety check: Pass | BSD-3-Clause | 1 mo ago |
| 21 | 根据设计文档生成 AscendC 算子完整代码实现并完成框架适配。TRIGGER when: 设计文档已完成,需要生成 ophost/opkernel 代码、注册到 PyTorch 框架、编译测试。关键词:代码生成、ophost、opkernel、tiling、kernel、框架适配、算子注册。 | ascend-ai-coding/ | 174 | — | ~2k | Automated safety check: Pass | No licence | today |