Topic · AI & LLM Engineering
Best computer vision skills, page 3
Computer vision skills, ranked
Ranked by score. Sort bymost stars,trending,newest,recently updated
| # | Skill | Repository | Stars | Used in | Tokens | Auto-check | Licence | Updated |
|---|---|---|---|---|---|---|---|---|
| 97 | A skill your agent uses to bring a supported object-detection vision model from HuggingFace or NVIDIA NGC into an NVIDIA DeepStream pipeline with end-to-end automation: ONNX download, SafeTensors… | NVIDIA/ | 3.5k | — | ~3.6k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 98 | Sparse4D for multi-camera temporal 3D object detection and tracking. | NVIDIA/ | 3.5k | — | ~3.8k | Automated safety check: Notes | Apache-2.0 | yesterday |
| 99 | Visual ChangeNet for binary image classification and segmentation in AOI defect detection. | NVIDIA/ | 3.5k | — | ~4.7k | Automated safety check: Notes | Apache-2.0 | yesterday |
| 100 | 100.Fal Vision Analyze images — segment objects, detect, run OCR, describe, and answer visual questions via fal.ai vision models. | nexu-io/ | 100k | — | ~295 | Automated safety check: Pass | Apache-2.0 | yesterday |
| 101 | Develop and debug the Multimodal DataPrep microservice: its FastAPI media endpoints, in-process embedding pipeline, batch jobs, object detection, telemetry and Metrics Manager publishing, and… | open-edge-platform/ | 168 | — | ~1.3k | Automated safety check: Pass | Apache-2.0 | today |
| 102 | 102.Kaiming He Applies the reasoning style of Kaiming He, computer vision pioneer and creator of ResNet. | K-Dense-AI/ | 282 | — | ~1.6k | Automated safety check: Pass | MIT | 1 mo ago |
| 103 | OpenSpec-mode authoring for Chorus PM workflows in Kiro CLI. | Chorus-AIDLC/ | 1.2k | — | ~6.8k | Automated safety check: Pass | AGPL-3.0 | today |
| 104 | 104.Chorus Yolo Full-auto AI-DLC pipeline — from prompt to done. An agent skill from Chorus-AIDLC/Chorus. | Chorus-AIDLC/ | 1.2k | — | ~6.7k | Automated safety check: Pass | AGPL-3.0 | today |
| 105 | 105.Openspec Aware OpenSpec-mode authoring for Chorus PM workflows in Hermes. An agent skill from Chorus-AIDLC/Chorus. | Chorus-AIDLC/ | 1.2k | — | ~7.2k | Automated safety check: Notes | AGPL-3.0 | today |
| 106 | 106.Openspec Aware OpenSpec-mode authoring for Chorus PM workflows in Pi. An agent skill from Chorus-AIDLC/Chorus. | Chorus-AIDLC/ | 1.2k | — | ~7.2k | Automated safety check: Pass | AGPL-3.0 | today |
| 107 | 107.Openspec Aware OpenSpec-mode authoring for Chorus PM workflows on OpenClaw — the default whenever OpenSpec is usable. | Chorus-AIDLC/ | 1.2k | — | ~7.6k | Automated safety check: Pass | AGPL-3.0 | today |
| 108 | 108.Openspec Aware OpenSpec-mode authoring for Chorus PM workflows in Codex. An agent skill from Chorus-AIDLC/Chorus. | Chorus-AIDLC/ | 1.2k | — | ~7.3k | Automated safety check: Pass | AGPL-3.0 | today |
| 109 | 109.Openspec Aware OpenSpec-mode authoring for Chorus PM workflows in Claude Code. | Chorus-AIDLC/ | 1.2k | — | ~6.9k | Automated safety check: Pass | AGPL-3.0 | today |
| 110 | OpenSpec-mode authoring for Chorus PM workflows on dsh — the default whenever OpenSpec is usable. | Chorus-AIDLC/ | 1.2k | — | ~7.5k | Automated safety check: Pass | AGPL-3.0 | today |
| 111 | 111.Yolo Full-auto AI-DLC pipeline in Hermes — from prompt to done. An agent skill from Chorus-AIDLC/Chorus. | Chorus-AIDLC/ | 1.2k | — | ~8.5k | Automated safety check: Pass | AGPL-3.0 | today |
| 112 | 112.Yolo Full-auto AI-DLC pipeline — from prompt to done. An agent skill from Chorus-AIDLC/Chorus. | Chorus-AIDLC/ | 1.2k | — | ~6.8k | Automated safety check: Pass | AGPL-3.0 | today |
| 113 | 113.Yolo Chorus Full-auto AI-DLC pipeline — drive a single prompt from Idea through Proposal, Execution, and Verification to Done. | Chorus-AIDLC/ | 1.2k | — | ~7.8k | Automated safety check: Pass | AGPL-3.0 | today |
| 114 | End-to-end ONNX PTQ workflow for AMD Quark — from a .onnx file (and calibration data) to a quantized .onnx output. | amd/ | 181 | — | ~4.5k | Automated safety check: Pass | MIT | 10 days ago |
| 115 | Operate and troubleshoot the UAV Mission Compute SDK for PX4 telemetry, camera streaming, missions, computer vision, and edge AI demonstrations. | open-edge-platform/ | 140 | — | ~1.1k | Automated safety check: Notes | Apache-2.0 | yesterday |
| 116 | SOTA Computer Vision Expert (2026). An agent skill from aiskillstore/marketplace. | aiskillstore/ | 430 | 6 repos | ~971 | Automated safety check: Pass | No licence | today |
| 117 | NVIDIA DeepStream SDK development with Python pyservicemaker API. | NVIDIA/ | 3.5k | — | ~3.3k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 118 | A skill your agent uses when building, deploying, evaluating, debugging, or measuring latency for the DeepStream SOP Inference Microservice — a GPU-accelerated FastAPI service that detects whether… | NVIDIA/ | 3.5k | — | ~4.7k | Automated safety check: Notes | Apache-2.0 | yesterday |
| 119 | CLIP vision-language model for image-text retrieval, zero-shot classification, embedding extraction, ONNX export, and TensorRT deployment. | NVIDIA/ | 3.5k | — | ~4k | Automated safety check: Notes | Apache-2.0 | yesterday |
| 120 | Fine-tune any HuggingFace CV / VLM / LLM model on local NVIDIA GPUs inside an NGC PyTorch container when no dedicated TAO model skill matches. | NVIDIA/ | 3.5k | — | ~4.9k | Automated safety check: Notes | Apache-2.0 | yesterday |
| 121 | BEVFusion for multi-sensor 3D object detection. An agent skill from NVIDIA/skills. | NVIDIA/ | 3.5k | — | ~3.4k | Automated safety check: Notes | Apache-2.0 | yesterday |
| 122 | Co-DETR (CoDINO) for object detection. An agent skill from NVIDIA/skills. | NVIDIA/ | 3.5k | — | ~4.8k | Automated safety check: Notes | Apache-2.0 | yesterday |
| 123 | Deformable DETR for 2D object detection. An agent skill from NVIDIA/skills. | NVIDIA/ | 3.5k | — | ~3.9k | Automated safety check: Notes | Apache-2.0 | yesterday |
| 124 | DINO (DETR with Improved DeNoising Anchor Boxes) for 2D object detection. | NVIDIA/ | 3.5k | — | ~2.8k | Automated safety check: Notes | Apache-2.0 | yesterday |
| 125 | Grounding DINO for open-set object detection. An agent skill from NVIDIA/skills. | NVIDIA/ | 3.5k | — | ~3.9k | Automated safety check: Notes | Apache-2.0 | yesterday |
| 126 | PyTorch-based TAO image classification. An agent skill from NVIDIA/skills. | NVIDIA/ | 3.5k | — | ~3.6k | Automated safety check: Notes | Apache-2.0 | yesterday |
| 127 | PointPillars for 3D object detection from LiDAR point clouds. | NVIDIA/ | 3.5k | — | ~4k | Automated safety check: Notes | Apache-2.0 | yesterday |
| 128 | RT-DETR (Real-Time DEtection TRansformer) for 2D object detection. | NVIDIA/ | 3.5k | — | ~4.5k | Automated safety check: Notes | Apache-2.0 | yesterday |
| 129 | Audit and improve an entire agentic-development setup, including skills, context, memory, tools, models, permissions, hooks, workflow habits, and secrets hygiene. | data-goblin/ | 1k | — | ~4.4k | Automated safety check: Notes | GPL-3.0 | 2 days ago |
| 130 | 130.Vlm Segmentation Choose and evaluate VLM or segmentation pipelines, including text-conditioned detection, masks, part labels, model-license constraints, and measured GPU deployment choices. | AnastasiyaW/ | 154 | — | ~910 | Automated safety check: Pass | MIT | 5 days ago |
| 131 | Route Quark ONNX user goals to the correct atomic skill. An agent skill from amd/Quark. | amd/ | 181 | — | ~2.7k | Automated safety check: Pass | MIT | 10 days ago |
| 132 | 132.API Media Call nodetool.media or nodetool.generations from a code action: generate or edit images, video, speech and music with a picked model, transcribe and embed, judge images with a vision model, read a… | nodetool-ai/ | 554 | — | ~1.9k | Automated safety check: Pass | AGPL-3.0 | today |
| 133 | Work with hyperspectral and multispectral images in MATLAB. An agent skill from matlab/matlab-agentic-toolkit. | matlab/ | 1.1k | — | ~3.7k | Automated safety check: Pass | Unknown | 7 days ago |
| 134 | Creates MATLAB interfaces to Python image processing and computer vision models from GitHub repositories or pip-installable packages using MPyReq. | matlab/ | 1.1k | — | ~3.7k | Automated safety check: Pass | Unknown | 7 days ago |
| 135 | Build, import, analyze, tolerate, and optimize optical systems using the Optical Design and Simulation Library. | matlab/ | 1.1k | — | ~3.8k | Automated safety check: Pass | Unknown | 7 days ago |
| 136 | Load this first for any task involving images, pictures, photos, scans, frames, volumes, or visual data — including reading, writing, filtering, enhancing, denoising, sharpening, deblurring… | matlab/ | 1.1k | — | ~3.7k | Automated safety check: Pass | Unknown | 7 days ago |
| 137 | Patterns for using blockedImage to process large images, harness parallel compute for image processing, and write custom adapters. | matlab/ | 1.1k | — | ~3k | Automated safety check: Pass | Unknown | 7 days ago |
| 138 | Read and write 3-D point cloud data using Lidar Toolbox file I/O. | matlab/ | 1.1k | — | ~3.5k | Automated safety check: Pass | Unknown | 7 days ago |
| 139 | A skill your agent uses to ask the VSS agent's videounderstanding tool a fresh visual question about a recorded clip. | NVIDIA/ | 3.5k | — | ~1.4k | Automated safety check: Notes | Apache-2.0 | yesterday |
| 140 | Expert in 3D computer vision labeling tools, workflows, and AI-assisted annotation for LiDAR, point clouds, and sensor fusion. | curiositech/ | 243 | 1 repo | ~3.2k | Automated safety check: Notes | MIT | 1 mo ago |
| 141 | 141.Drone Cv Expert Expert in drone systems, computer vision, and autonomous navigation. | curiositech/ | 243 | 1 repo | ~2.1k | Automated safety check: Pass | MIT | 1 mo ago |
| 142 | Advanced CV for infrastructure inspection including forest fire detection, wildfire precondition assessment, roof inspection, hail damage analysis, thermal imaging, and 3D Gaussian Splatting… | curiositech/ | 243 | 1 repo | ~2.2k | Automated safety check: Pass | MIT | 1 mo ago |
| 143 | Integrate a HuggingFace Computer Vision model into the NVIDIA TAO Toolkit ecosystem (tao-core config, tao-pytorch trainer, tao-deploy TensorRT pipeline). | NVIDIA/ | 3.5k | — | ~4.5k | Automated safety check: Notes | Apache-2.0 | yesterday |
| 144 | Based on computer vision, analyzes pet health indicators such as feeding frequency, drinking frequency, excretion status, mental state, vomiting behavior, and limping abnormalities through… | LeoYeAI/ | 2.2k | — | ~1.6k | Automated safety check: Pass | MIT | 2 mo ago |
Explore related skills
Category
More topics in AI & LLM Engineering
- Building AI agents525
- Deep learning408
- Embeddings381
- LLM inference and serving364
- Retrieval-augmented generation360
- Prompt engineering350
- Fine-tuning313
- LLM evaluation303
- Speech recognition and synthesis272
- Structured output and tool calling271
- LLM cost and token optimization256
- LLM API integration218
- LLM observability217
- Model routing and gateways217
- LLM guardrails208
- Model hubs and datasets180
- GPU and accelerator computing171
- Diffusion and image models167
- Natural language processing141
- Reinforcement learning67
- AI interpretability23