Search
Computer vision
Skills
Sort:BestMost starsTrending todayTrending this weekTrending this monthNewestRecently updatedName
| # | Skill | Repository | Stars | Used in | Tokens | Auto-check | Licence | Updated |
|---|---|---|---|---|---|---|---|---|
| 97 | Visual ChangeNet for binary image classification and segmentation in AOI defect detection. | NVIDIA/ | 3.6k | — | ~4.7k | Automated safety check: Notes | Apache-2.0 | yesterday |
| 98 | 98.Fal Vision Analyze images — segment objects, detect, run OCR, describe, and answer visual questions via fal.ai vision models. | nexu-io/ | 100k | — | ~295 | Automated safety check: Pass | Apache-2.0 | today |
| 99 | Develop and debug the Multimodal DataPrep microservice: its FastAPI media endpoints, in-process embedding pipeline, batch jobs, object detection, telemetry and Metrics Manager publishing, and… | open-edge-platform/ | 171 | — | ~1.3k | Automated safety check: Pass | Apache-2.0 | today |
| 100 | 100.Kaiming He Applies the reasoning style of Kaiming He, computer vision pioneer and creator of ResNet. | K-Dense-AI/ | 282 | — | ~1.6k | Automated safety check: Pass | MIT | 1 mo ago |
| 101 | OpenSpec-mode authoring for Chorus PM workflows in Kiro CLI. | Chorus-AIDLC/ | 1.2k | — | ~6.8k | Automated safety check: Pass | AGPL-3.0 | yesterday |
| 102 | 102.Chorus Yolo Full-auto AI-DLC pipeline — from prompt to done. An agent skill from Chorus-AIDLC/Chorus. | Chorus-AIDLC/ | 1.2k | — | ~6.7k | Automated safety check: Pass | AGPL-3.0 | yesterday |
| 103 | 103.Openspec Aware OpenSpec-mode authoring for Chorus PM workflows in Hermes. An agent skill from Chorus-AIDLC/Chorus. | Chorus-AIDLC/ | 1.2k | — | ~7.2k | Automated safety check: Notes | AGPL-3.0 | yesterday |
| 104 | 104.Openspec Aware OpenSpec-mode authoring for Chorus PM workflows in Pi. An agent skill from Chorus-AIDLC/Chorus. | Chorus-AIDLC/ | 1.2k | — | ~7.2k | Automated safety check: Pass | AGPL-3.0 | yesterday |
| 105 | 105.Openspec Aware OpenSpec-mode authoring for Chorus PM workflows on OpenClaw — the default whenever OpenSpec is usable. | Chorus-AIDLC/ | 1.2k | — | ~7.6k | Automated safety check: Pass | AGPL-3.0 | yesterday |
| 106 | 106.Openspec Aware OpenSpec-mode authoring for Chorus PM workflows in Codex. An agent skill from Chorus-AIDLC/Chorus. | Chorus-AIDLC/ | 1.2k | — | ~7.3k | Automated safety check: Pass | AGPL-3.0 | yesterday |
| 107 | 107.Openspec Aware OpenSpec-mode authoring for Chorus PM workflows in Claude Code. | Chorus-AIDLC/ | 1.2k | — | ~6.9k | Automated safety check: Pass | AGPL-3.0 | yesterday |
| 108 | OpenSpec-mode authoring for Chorus PM workflows on dsh — the default whenever OpenSpec is usable. | Chorus-AIDLC/ | 1.2k | — | ~7.5k | Automated safety check: Pass | AGPL-3.0 | yesterday |
| 109 | 109.Yolo Full-auto AI-DLC pipeline in Hermes — from prompt to done. An agent skill from Chorus-AIDLC/Chorus. | Chorus-AIDLC/ | 1.2k | — | ~8.5k | Automated safety check: Pass | AGPL-3.0 | yesterday |
| 110 | 110.Yolo Full-auto AI-DLC pipeline — from prompt to done. An agent skill from Chorus-AIDLC/Chorus. | Chorus-AIDLC/ | 1.2k | — | ~6.8k | Automated safety check: Pass | AGPL-3.0 | yesterday |
| 111 | 111.Yolo Chorus Full-auto AI-DLC pipeline — drive a single prompt from Idea through Proposal, Execution, and Verification to Done. | Chorus-AIDLC/ | 1.2k | — | ~7.8k | Automated safety check: Pass | AGPL-3.0 | yesterday |
| 112 | End-to-end ONNX PTQ workflow for AMD Quark — from a .onnx file (and calibration data) to a quantized .onnx output. | amd/ | 182 | — | ~4.5k | Automated safety check: Pass | MIT | 12 days ago |
| 113 | Operate and troubleshoot the UAV Mission Compute SDK for PX4 telemetry, camera streaming, missions, computer vision, and edge AI demonstrations. | open-edge-platform/ | 140 | — | ~1.1k | Automated safety check: Notes | Apache-2.0 | yesterday |
| 114 | Process images using object detection, classification, and segmentation. | jeremylongshore/ | 2.8k | — | ~898 | Automated safety check: Pass | MIT | today |
| 115 | Audit and improve an entire agentic-development setup, including skills, context, memory, tools, models, permissions, hooks, workflow habits, and secrets hygiene. | data-goblin/ | 1k | — | ~4.4k | Automated safety check: Notes | GPL-3.0 | 3 days ago |
| 116 | NVIDIA DeepStream SDK development with Python pyservicemaker API. | NVIDIA/ | 3.6k | — | ~3.3k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 117 | A skill your agent uses when building, deploying, evaluating, debugging, or measuring latency for the DeepStream SOP Inference Microservice — a GPU-accelerated FastAPI service that detects whether… | NVIDIA/ | 3.6k | — | ~4.7k | Automated safety check: Notes | Apache-2.0 | yesterday |
| 118 | CLIP vision-language model for image-text retrieval, zero-shot classification, embedding extraction, ONNX export, and TensorRT deployment. | NVIDIA/ | 3.6k | — | ~4k | Automated safety check: Notes | Apache-2.0 | yesterday |
| 119 | Fine-tune any HuggingFace CV / VLM / LLM model on local NVIDIA GPUs inside an NGC PyTorch container when no dedicated TAO model skill matches. | NVIDIA/ | 3.6k | — | ~4.9k | Automated safety check: Notes | Apache-2.0 | yesterday |
| 120 | BEVFusion for multi-sensor 3D object detection. An agent skill from NVIDIA/skills. | NVIDIA/ | 3.6k | — | ~3.4k | Automated safety check: Notes | Apache-2.0 | yesterday |
| 121 | Co-DETR (CoDINO) for object detection. An agent skill from NVIDIA/skills. | NVIDIA/ | 3.6k | — | ~4.8k | Automated safety check: Notes | Apache-2.0 | yesterday |
| 122 | Deformable DETR for 2D object detection. An agent skill from NVIDIA/skills. | NVIDIA/ | 3.6k | — | ~3.9k | Automated safety check: Notes | Apache-2.0 | yesterday |
| 123 | DINO (DETR with Improved DeNoising Anchor Boxes) for 2D object detection. | NVIDIA/ | 3.6k | — | ~2.8k | Automated safety check: Notes | Apache-2.0 | yesterday |
| 124 | Grounding DINO for open-set object detection. An agent skill from NVIDIA/skills. | NVIDIA/ | 3.6k | — | ~3.9k | Automated safety check: Notes | Apache-2.0 | yesterday |
| 125 | PyTorch-based TAO image classification. An agent skill from NVIDIA/skills. | NVIDIA/ | 3.6k | — | ~3.6k | Automated safety check: Notes | Apache-2.0 | yesterday |
| 126 | PointPillars for 3D object detection from LiDAR point clouds. | NVIDIA/ | 3.6k | — | ~4k | Automated safety check: Notes | Apache-2.0 | yesterday |
| 127 | RT-DETR (Real-Time DEtection TRansformer) for 2D object detection. | NVIDIA/ | 3.6k | — | ~4.5k | Automated safety check: Notes | Apache-2.0 | yesterday |
| 128 | 128.Vlm Segmentation Choose and evaluate VLM or segmentation pipelines, including text-conditioned detection, masks, part labels, model-license constraints, and measured GPU deployment choices. | AnastasiyaW/ | 154 | — | ~910 | Automated safety check: Pass | MIT | yesterday |
| 129 | SOTA Computer Vision Expert (2026). An agent skill from aiskillstore/marketplace. | aiskillstore/ | 433 | 5 repos | ~971 | Automated safety check: Pass | No licence | today |
| 130 | Meta Quest Passthrough Camera Access (PCA) for Unity — access the forward-facing RGB cameras on Quest 3 / Quest 3S to feed Computer Vision and Machine Learning pipelines. | meta-quest/ | 215 | — | ~1.6k | Automated safety check: Pass | Apache-2.0 | 16 days ago |
| 131 | Route Quark ONNX user goals to the correct atomic skill. An agent skill from amd/Quark. | amd/ | 182 | — | ~2.7k | Automated safety check: Pass | MIT | 12 days ago |
| 132 | 132.API Media Call nodetool.media or nodetool.generations from a code action: generate or edit images, video, speech and music with a picked model, transcribe and embed, judge images with a vision model, read a… | nodetool-ai/ | 560 | — | ~1.9k | Automated safety check: Pass | AGPL-3.0 | today |
| 133 | Work with hyperspectral and multispectral images in MATLAB. An agent skill from matlab/matlab-agentic-toolkit. | matlab/ | 1.1k | — | ~3.7k | Automated safety check: Pass | Unknown | 2 days ago |
| 134 | Creates MATLAB interfaces to Python image processing and computer vision models from GitHub repositories or pip-installable packages using MPyReq. | matlab/ | 1.1k | — | ~3.7k | Automated safety check: Pass | Unknown | 2 days ago |
| 135 | Build, import, analyze, tolerate, and optimize optical systems using the Optical Design and Simulation Library. | matlab/ | 1.1k | — | ~3.8k | Automated safety check: Pass | Unknown | 2 days ago |
| 136 | Load this first for any task involving images, pictures, photos, scans, frames, volumes, or visual data — including reading, writing, filtering, enhancing, denoising, sharpening, deblurring… | matlab/ | 1.1k | — | ~3.7k | Automated safety check: Pass | Unknown | 2 days ago |
| 137 | Patterns for using blockedImage to process large images, harness parallel compute for image processing, and write custom adapters. | matlab/ | 1.1k | — | ~3k | Automated safety check: Pass | Unknown | 2 days ago |
| 138 | Read and write 3-D point cloud data using Lidar Toolbox file I/O. | matlab/ | 1.1k | — | ~3.5k | Automated safety check: Pass | Unknown | 2 days ago |
| 139 | Display images and annotations for image processing, computer vision, and visual inspection. | matlab/ | 1.1k | — | ~3.8k | Automated safety check: Pass | Unknown | 2 days ago |
| 140 | A skill your agent uses to ask the VSS agent's videounderstanding tool a fresh visual question about a recorded clip. | NVIDIA/ | 3.6k | — | ~1.4k | Automated safety check: Notes | Apache-2.0 | yesterday |
| 141 | Analyze videos using Google's Gemini API - describe content, answer questions, transcribe audio with visual descriptions, reference timestamps, clip videos, and process YouTube URLs. | einverne/ | 121 | — | ~2.6k | Automated safety check: Notes | MIT | 1 mo ago |
| 142 | Integrate a HuggingFace Computer Vision model into the NVIDIA TAO Toolkit ecosystem (tao-core config, tao-pytorch trainer, tao-deploy TensorRT pipeline). | NVIDIA/ | 3.6k | — | ~4.5k | Automated safety check: Notes | Apache-2.0 | yesterday |
| 143 | Based on computer vision, analyzes pet health indicators such as feeding frequency, drinking frequency, excretion status, mental state, vomiting behavior, and limping abnormalities through… | LeoYeAI/ | 2.2k | — | ~1.6k | Automated safety check: Pass | MIT | 2 mo ago |
| 144 | 144.Core ML Core ML, Create ML, Vision framework, Natural Language framework, on-device ML integration. | gustavscirulis/ | 116 | 1 repo | ~3.8k | Automated safety check: Notes | Unknown | 5 mo ago |