Search

ONNX · LLM inference and serving

26 skills found.
Search results
#SkillRepositoryStarsUsed inTokensAuto-checkLicenceUpdated
1

Install, convert, debug, and benchmark sim2real ONNX GPU and TensorRT inference backends on onboard JetPack 5 Orin hosts such as g1-cable.

EGalahad/sim2real146—~1.1kAutomated safety check: PassNo licence13 days ago
2

AIPC, AI Porting Conversion. An agent skill from qualcomm/qai-appbuilder.

qualcomm/qai-appbuilder247—~5.7kAutomated safety check: NotesUnknownyesterday
3

QAI ModelBuilder. An agent skill from qualcomm/qai-appbuilder.

qualcomm/qai-appbuilder247—~4.1kAutomated safety check: PassBSD-3-Clauseyesterday
4

Explains where HOT-Step generation time goes (LM/DiT/VAE), how the TensorRT paths activate, how to benchmark from logs, and which knobs trade quality for speed.

scragnog/HOT-Step-CPP174—~4.9kAutomated safety check: PassMIT2 days ago
5

Run, resume, monitor, diagnose, and report Quark Quant-Perf workflows for PyTorch and HuggingFace transformers models.

amd/Quark182—~3kAutomated safety check: PassMIT13 days ago
6

Author a new ShapeShifter graph-transformation pass for AMD Quark (ONNX or PyTorch) so it conforms to the pass framework's conventions and auto-registers.

amd/Quark182—~2.9kAutomated safety check: PassMIT13 days ago
7

L3 recipe that runs quark.onnx.AutoSearchPro end-to-end on a user .onnx model: intake → preset selection (or custom search space) → calibration / eval data reader → standalone autosearch script…

amd/Quark182—~3.4kAutomated safety check: PassMIT13 days ago
8

Diagnose failed Quark ONNX installation, calibration, quantization, custom-op compilation, or export attempts.

amd/Quark182—~4.8kAutomated safety check: PassMIT13 days ago
9

Run OpenMed models fully on-device with the MLX (Apple Silicon), CoreML (iOS/macOS), or ONNX/WebGPU (cross-platform/browser) backends, including convert-quantize-run workflows.

maziyarpanahi/openmed5.5k—~2kAutomated safety check: PassApache-2.0today
10

Inspect a target ONNX model and prepare metadata for Quark ONNX PTQ planning.

amd/Quark182—~4.3kAutomated safety check: PassMIT13 days ago
11

End-to-end ONNX PTQ workflow for AMD Quark — from a .onnx file (and calibration data) to a quantized .onnx output.

amd/Quark182—~4.5kAutomated safety check: PassMIT13 days ago
12

CLIP vision-language model for image-text retrieval, zero-shot classification, embedding extraction, ONNX export, and TensorRT deployment.

NVIDIA/skills3.6k—~4kAutomated safety check: NotesApache-2.02 days ago
13

InternVideo2-CLIP L14 (TAO videoclip) for video-text retrieval, zero-shot classification, embedding extraction, LoRA fine-tuning, ONNX export, and TensorRT deployment.

NVIDIA/skills3.6k—~3.5kAutomated safety check: NotesApache-2.02 days ago
14

Build a Quark ONNX PTQ quantization plan from modelanalysis.json and user intent.

amd/Quark182—~4.8kAutomated safety check: PassMIT13 days ago
15

Validate Quark ONNX quantization output using four lightweight checks: auxiliary file copy alignment, expected non-quantized initializer MD5 byte-identity (inline rawdata + external-data byte…

amd/Quark182—~2.5kAutomated safety check: PassMIT13 days ago
16

Route Quark ONNX user goals to the correct atomic skill. An agent skill from amd/Quark.

amd/Quark182—~2.7kAutomated safety check: PassMIT13 days ago
17

Apply existing ShapeShifter graph passes to an .onnx model via the quark-cli shapeshifter CLI or a ShapeShifter YAML.

amd/Quark182—~1.9kAutomated safety check: PassMIT13 days ago
18

Detect upstream Quark ONNX changes that affect the ONNX skill family and classify required updates.

amd/Quark182—~3.3kAutomated safety check: PassMIT13 days ago
19

Partition an ONNX model graph into named functional subgraphs and emit a subgraphpartition.json file.

amd/Quark182—~2.7kAutomated safety check: PassMIT13 days ago
20

Prepare export and downstream evaluation handoff for a planned or completed Quark PTQ run.

amd/Quark182—~1.5kAutomated safety check: PassMIT13 days ago
21

L3 recipe that runs a Torch LLM PTQ end-to-end for AMD Quark — for PyTorch / HuggingFace transformers models (safetensors input): quantize → validate → evaluate.

amd/Quark182—~2.6kAutomated safety check: PassMIT13 days ago
22

Installs or verifies AMD Quark and ensures the selected Python environment has an accelerator-matched PyTorch.

amd/Quark182—~3.5kAutomated safety check: NotesMIT13 days ago
23

Runs an end-to-end AMD Quark post-training quantization workflow for PyTorch / Hugging Face LLMs: inspect a Hub or local model, choose a quantization plan, create reproducible artifacts, request…

amd/Quark182—~2.3kAutomated safety check: PassMIT13 days ago
24

Integrate a HuggingFace Computer Vision model into the NVIDIA TAO Toolkit ecosystem (tao-core config, tao-pytorch trainer, tao-deploy TensorRT pipeline).

NVIDIA/skills3.6k—~4.5kAutomated safety check: NotesApache-2.02 days ago
25

优化实际模型推理链路,将正确性对齐、分段 profiling、显存与数据搬运、TensorRT/ONNX/PyTorch 后端、attention/kernel、FP8/compile、缓存与少步采样、质量回归、GPU 成本和服务验收串成同一实验闭环。当用户要求推理提速、降低显存或 GPU 成本、复现模型效果、定位 GPU 利用率低、优化图像/视频/扩散模型或自托管 LLM 时使用,提供…

majiayu000/spellbook287—~1.1kAutomated safety check: PassMIT3 days ago
26

Build machine vision inspection systems with MATLAB Visual Inspection Toolbox.

matlab/matlab-agentic-toolkit1.1k—~3.1kAutomated safety check: PassUnknown3 days ago