Search
ONNX · LLM inference and serving
Skills
Sort:BestMost starsTrending todayTrending this weekTrending this monthNewestRecently updatedName
| # | Skill | Repository | Stars | Used in | Tokens | Auto-check | Licence | Updated |
|---|---|---|---|---|---|---|---|---|
| 1 | Install, convert, debug, and benchmark sim2real ONNX GPU and TensorRT inference backends on onboard JetPack 5 Orin hosts such as g1-cable. | EGalahad/ | 146 | — | ~1.1k | Automated safety check: Pass | No licence | 13 days ago |
| 2 | AIPC, AI Porting Conversion. An agent skill from qualcomm/qai-appbuilder. | qualcomm/ | 247 | — | ~5.7k | Automated safety check: Notes | Unknown | yesterday |
| 3 | QAI ModelBuilder. An agent skill from qualcomm/qai-appbuilder. | qualcomm/ | 247 | — | ~4.1k | Automated safety check: Pass | BSD-3-Clause | yesterday |
| 4 | Explains where HOT-Step generation time goes (LM/DiT/VAE), how the TensorRT paths activate, how to benchmark from logs, and which knobs trade quality for speed. | scragnog/ | 174 | — | ~4.9k | Automated safety check: Pass | MIT | 2 days ago |
| 5 | Run, resume, monitor, diagnose, and report Quark Quant-Perf workflows for PyTorch and HuggingFace transformers models. | amd/ | 182 | — | ~3k | Automated safety check: Pass | MIT | 13 days ago |
| 6 | Author a new ShapeShifter graph-transformation pass for AMD Quark (ONNX or PyTorch) so it conforms to the pass framework's conventions and auto-registers. | amd/ | 182 | — | ~2.9k | Automated safety check: Pass | MIT | 13 days ago |
| 7 | L3 recipe that runs quark.onnx.AutoSearchPro end-to-end on a user .onnx model: intake → preset selection (or custom search space) → calibration / eval data reader → standalone autosearch script… | amd/ | 182 | — | ~3.4k | Automated safety check: Pass | MIT | 13 days ago |
| 8 | Diagnose failed Quark ONNX installation, calibration, quantization, custom-op compilation, or export attempts. | amd/ | 182 | — | ~4.8k | Automated safety check: Pass | MIT | 13 days ago |
| 9 | Run OpenMed models fully on-device with the MLX (Apple Silicon), CoreML (iOS/macOS), or ONNX/WebGPU (cross-platform/browser) backends, including convert-quantize-run workflows. | maziyarpanahi/ | 5.5k | — | ~2k | Automated safety check: Pass | Apache-2.0 | today |
| 10 | Inspect a target ONNX model and prepare metadata for Quark ONNX PTQ planning. | amd/ | 182 | — | ~4.3k | Automated safety check: Pass | MIT | 13 days ago |
| 11 | End-to-end ONNX PTQ workflow for AMD Quark — from a .onnx file (and calibration data) to a quantized .onnx output. | amd/ | 182 | — | ~4.5k | Automated safety check: Pass | MIT | 13 days ago |
| 12 | CLIP vision-language model for image-text retrieval, zero-shot classification, embedding extraction, ONNX export, and TensorRT deployment. | NVIDIA/ | 3.6k | — | ~4k | Automated safety check: Notes | Apache-2.0 | 2 days ago |
| 13 | InternVideo2-CLIP L14 (TAO videoclip) for video-text retrieval, zero-shot classification, embedding extraction, LoRA fine-tuning, ONNX export, and TensorRT deployment. | NVIDIA/ | 3.6k | — | ~3.5k | Automated safety check: Notes | Apache-2.0 | 2 days ago |
| 14 | Build a Quark ONNX PTQ quantization plan from modelanalysis.json and user intent. | amd/ | 182 | — | ~4.8k | Automated safety check: Pass | MIT | 13 days ago |
| 15 | Validate Quark ONNX quantization output using four lightweight checks: auxiliary file copy alignment, expected non-quantized initializer MD5 byte-identity (inline rawdata + external-data byte… | amd/ | 182 | — | ~2.5k | Automated safety check: Pass | MIT | 13 days ago |
| 16 | Route Quark ONNX user goals to the correct atomic skill. An agent skill from amd/Quark. | amd/ | 182 | — | ~2.7k | Automated safety check: Pass | MIT | 13 days ago |
| 17 | Apply existing ShapeShifter graph passes to an .onnx model via the quark-cli shapeshifter CLI or a ShapeShifter YAML. | amd/ | 182 | — | ~1.9k | Automated safety check: Pass | MIT | 13 days ago |
| 18 | Detect upstream Quark ONNX changes that affect the ONNX skill family and classify required updates. | amd/ | 182 | — | ~3.3k | Automated safety check: Pass | MIT | 13 days ago |
| 19 | Partition an ONNX model graph into named functional subgraphs and emit a subgraphpartition.json file. | amd/ | 182 | — | ~2.7k | Automated safety check: Pass | MIT | 13 days ago |
| 20 | Prepare export and downstream evaluation handoff for a planned or completed Quark PTQ run. | amd/ | 182 | — | ~1.5k | Automated safety check: Pass | MIT | 13 days ago |
| 21 | L3 recipe that runs a Torch LLM PTQ end-to-end for AMD Quark — for PyTorch / HuggingFace transformers models (safetensors input): quantize → validate → evaluate. | amd/ | 182 | — | ~2.6k | Automated safety check: Pass | MIT | 13 days ago |
| 22 | Installs or verifies AMD Quark and ensures the selected Python environment has an accelerator-matched PyTorch. | amd/ | 182 | — | ~3.5k | Automated safety check: Notes | MIT | 13 days ago |
| 23 | Runs an end-to-end AMD Quark post-training quantization workflow for PyTorch / Hugging Face LLMs: inspect a Hub or local model, choose a quantization plan, create reproducible artifacts, request… | amd/ | 182 | — | ~2.3k | Automated safety check: Pass | MIT | 13 days ago |
| 24 | Integrate a HuggingFace Computer Vision model into the NVIDIA TAO Toolkit ecosystem (tao-core config, tao-pytorch trainer, tao-deploy TensorRT pipeline). | NVIDIA/ | 3.6k | — | ~4.5k | Automated safety check: Notes | Apache-2.0 | 2 days ago |
| 25 | 优化实际模型推理链路,将正确性对齐、分段 profiling、显存与数据搬运、TensorRT/ONNX/PyTorch 后端、attention/kernel、FP8/compile、缓存与少步采样、质量回归、GPU 成本和服务验收串成同一实验闭环。当用户要求推理提速、降低显存或 GPU 成本、复现模型效果、定位 GPU 利用率低、优化图像/视频/扩散模型或自托管 LLM 时使用,提供… | majiayu000/ | 287 | — | ~1.1k | Automated safety check: Pass | MIT | 3 days ago |
| 26 | Build machine vision inspection systems with MATLAB Visual Inspection Toolbox. | matlab/ | 1.1k | — | ~3.1k | Automated safety check: Pass | Unknown | 3 days ago |