Search
Computer vision
Skills
Sort:BestMost starsTrending todayTrending this weekTrending this monthNewestRecently updatedName
| # | Skill | Repository | Stars | Used in | Tokens | Auto-check | Licence | Updated |
|---|---|---|---|---|---|---|---|---|
| 145 | Computer vision engineering for object detection, segmentation, and visual AI, covering CNN and Vision Transformer architectures and ONNX/TensorRT deployment. | borghei/ | 891 | — | ~1.8k | Automated safety check: Pass | MIT | 3 days ago |
| 146 | Display 3-D image volumes, medical image volumes, surface meshes, and annotations for 3-D image processing. | matlab/ | 1.1k | — | ~4.8k | Automated safety check: Pass | Unknown | 2 days ago |
| 147 | Register 3-D point clouds using ICP, NDT, LOAM, FGR, phase correlation, and CPD algorithms. | matlab/ | 1.1k | — | ~4.4k | Automated safety check: Pass | Unknown | 2 days ago |
| 148 | 148.Scholar Compute Design and execute computational social science analyses across 11 modules: text-as-data/NLP (STM, BERTopic, Wordfish, BERT, conText embedding regression, LLM annotation + DSL bias correction… | joshzyj/ | 168 | — | ~15k | Automated safety check: Pass | Unknown | 22 days ago |
| 149 | Computer vision for bio-image preprocessing, feature detection, real-time microscopy. | jaechang-hits/ | 374 | 1 repo | ~3.9k | Automated safety check: Pass | Apache-2.0 | 11 days ago |
| 150 | Deploy and maintain image understanding (OCR + local VLM + cloud VL) and image generation for Codex connected to text-only models like DeepSeek. | aiskillstore/ | 433 | — | ~814 | Automated safety check: Notes | MIT | yesterday |
| 151 | Research, write, code, analyze, present, and develop proposals for thermal-fluid mechanical engineering work with source-aware rigor. | ccplugins/ | 970 | — | ~2.9k | Automated safety check: Pass | Apache-2.0 | 1 mo ago |
| 152 | 152.Ops Business operations command center. An agent skill from davepoon/buildwithclaude. | davepoon/ | 3.6k | — | ~999 | Automated safety check: Notes | MIT | yesterday |
| 153 | 153.Ops Yolo YOLO mode. An agent skill from davepoon/buildwithclaude. | davepoon/ | 3.6k | — | ~2.9k | Automated safety check: Notes | MIT | yesterday |
| 154 | 154.Ffmpeg A skill your agent uses for local FFmpeg/FFprobe media inspection, remuxing, transcoding, filtering, evidence-bounded video review, transcript-assisted editorial plans, edit decision lists… | magnus919/ | 115 | — | ~3k | Automated safety check: Pass | MIT | today |
| 155 | A skill your agent uses when image reading, screenshot analysis, visual diff, design mockup analysis, pasted image handling, or Z.AI Vision MCP fails, times out, returns 429/Too Many Requests, or… | skuramatata/ | 114 | — | ~1k | Automated safety check: Pass | No licence | 3 mo ago |
| 156 | A skill your agent uses when targeting Asian Conference on Computer Vision (ACCV) or deciding whether a computer-science manuscript fits this venue. | franklee16/ | 223 | 1 repo | ~1.9k | Automated safety check: Pass | No licence | 22 days ago |
| 157 | A skill your agent uses when targeting British Machine Vision Conference (BMVC) or deciding whether a computer-science manuscript fits this venue. | franklee16/ | 223 | 1 repo | ~1.9k | Automated safety check: Pass | No licence | 22 days ago |
| 158 | A skill your agent uses when targeting IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) or deciding whether a computer-science manuscript fits this venue. | franklee16/ | 223 | 1 repo | ~2k | Automated safety check: Pass | No licence | 22 days ago |
| 159 | A skill your agent uses when targeting European Conference on Computer Vision (ECCV) or deciding whether a computer-science manuscript fits this venue. | franklee16/ | 223 | 1 repo | ~1.9k | Automated safety check: Pass | No licence | 22 days ago |
| 160 | A skill your agent uses when targeting IEEE Transactions on Pattern Analysis and Machine Intelligence (TPAMI) or deciding whether a computer vision or machine learning manuscript fits this archival… | franklee16/ | 223 | 1 repo | ~2k | Automated safety check: Pass | No licence | 22 days ago |
| 161 | A skill your agent uses when targeting IEEE/CVF International Conference on Computer Vision (ICCV) or deciding whether a computer-science manuscript fits this venue. | franklee16/ | 223 | 1 repo | ~1.9k | Automated safety check: Pass | No licence | 22 days ago |
| 162 | A skill your agent uses when targeting IEEE/CVF Winter Conference on Applications of Computer Vision (WACV) or deciding whether a computer-science manuscript fits this venue. | franklee16/ | 223 | 1 repo | ~2k | Automated safety check: Pass | No licence | 22 days ago |
| 163 | 163.Caffe Cifar 10 Guidance for building and training with the Caffe deep learning framework on CIFAR-10 dataset. | lazyFrogLOL/ | 128 | — | ~1.7k | Automated safety check: Pass | No licence | 4 mo ago |
| 164 | A skill your agent uses whenever the user wants to find, shortlist, vet, or enrich US AI/ML/data consulting firms (consultancies) — AI/ML development, MLOps, generative AI / LLM apps (RAG, chatbots… | jeremylongshore/ | 2.8k | — | ~4k | Automated safety check: Notes | MIT | yesterday |
| 165 | Apply computer vision research methods, models, and evaluation tools | wentorai/ | 298 | 1 repo | ~1.6k | Automated safety check: Pass | MIT | 3 mo ago |
| 166 | Guide to Transformer architectures for NLP and computer vision | wentorai/ | 298 | 1 repo | ~2.1k | Automated safety check: Pass | MIT | 3 mo ago |
| 167 | Build OCR pipelines in MATLAB using the ocr() function. An agent skill from matlab/matlab-agentic-toolkit. | matlab/ | 1.1k | — | ~5.2k | Automated safety check: Pass | Unknown | 2 days ago |
| 168 | Analyze videos with Google Gemini API (summaries, Q&A, transcription with timestamps + visual context, scene/timeline detection, video clipping, FPS control, multi-video comparison, and YouTube URL… | benchflow-ai/ | 1.8k | — | ~2.4k | Automated safety check: Pass | Apache-2.0 | 2 mo ago |
| 169 | 169.Openai Vision Analyze images and multi-frame sequences using OpenAI GPT vision models | benchflow-ai/ | 1.8k | — | ~5k | Automated safety check: Pass | Apache-2.0 | 2 mo ago |
| 170 | 170.Detection A skill your agent uses for PaddleViT object detection workflows with DETR, Swin, or PVTv2: validate COCO data, select configs, build/train/evaluate models, reason about transforms, losses… | VectorSpaceLab/ | 331 | — | ~2.4k | Automated safety check: Pass | Apache-2.0 | 1 mo ago |
| 171 | 171.Imgaug A skill your agent uses when working with imgaug image augmentation pipelines, aligned annotations, stochastic parameters, dtype/data utilities, or multicore augmentation for computer-vision data. | VectorSpaceLab/ | 331 | — | ~1.2k | Automated safety check: Pass | MIT | 1 mo ago |
| 172 | Use this sub-skill for Ultralytics YOLO predict workflows, source handling, streaming and batching, Results extraction, saving/plotting/cropping, and thread-safe inference. | VectorSpaceLab/ | 331 | — | ~1.4k | Automated safety check: Pass | AGPL-3.0 | 1 mo ago |
| 173 | 173.Mmdetection A skill your agent uses when working with MMDetection 3.x for object detection, instance/panoptic segmentation, tracking-adjacent configs, model zoo configs, inference, visualization… | VectorSpaceLab/ | 331 | — | ~1.2k | Automated safety check: Pass | Apache-2.0 | 1 mo ago |
| 174 | 174.Swin Transformer Use this repo skill for Microsoft Swin-Transformer image-classification model, config, data, checkpoint, SimMIM, Swin-MoE, and optional CUDA acceleration workflows. | VectorSpaceLab/ | 331 | — | ~1.2k | Automated safety check: Pass | MIT | 1 mo ago |
| 175 | A skill your agent uses for Ultralytics track mode, tracker YAML selection, persistent object IDs, and YOLO Solutions such as counting, heatmaps, speed, queue, region, similarity search, and… | VectorSpaceLab/ | 331 | — | ~871 | Automated safety check: Pass | AGPL-3.0 | 1 mo ago |
| 176 | 176.Ultralytics A skill your agent uses for Ultralytics YOLO package workflows: CLI/Python model usage, data/config setup, train/val, prediction/results, export/deployment, tracking/solutions, model-family… | VectorSpaceLab/ | 331 | — | ~1.2k | Automated safety check: Pass | AGPL-3.0 | 1 mo ago |
| 177 | Guide safe video-file, webcam, and optional half-precision demo use for pytorch-yolo-v3. | VectorSpaceLab/ | 331 | — | ~626 | Automated safety check: Pass | Apache-2.0 | 1 mo ago |
| 178 | 178.Yolov5 A skill your agent uses for clone-run Ultralytics YOLOv5 workflows: detection, segmentation, classification, export, benchmarks, datasets, weights, and Flask REST serving. | VectorSpaceLab/ | 331 | — | ~1.6k | Automated safety check: Pass | AGPL-3.0 | 1 mo ago |
| 179 | Paddle-based deep learning workflows from the course materials, including DNN/RNN text-style baselines and the CNN/LeNet image classification case using folder-labeled digit images. | Drchronx/ | 139 | — | ~516 | Automated safety check: Pass | Unknown | 4 mo ago |
| 180 | 180.AI Multimodal Process and generate multimedia content using Google Gemini API. | Microck/ | 404 | — | ~2.7k | Automated safety check: Notes | MIT | 1 mo ago |
| 181 | 181.Pp Twelvelabs Printing Press CLI for Twelvelabs. An agent skill from mvanhorn/printing-press-library. | mvanhorn/ | 2.1k | — | ~3.7k | Automated safety check: Notes | Apache-2.0 | yesterday |
| 182 | 182.Mindspeed Mm Vlm Universal VLM (vision-language understanding model) training guide for Huawei Ascend NPU using MindSpeed-MM. | ascend-ai-coding/ | 174 | — | ~5.2k | Automated safety check: Pass | No licence | today |
| 183 | 183.Vision See and understand images when you (the current model) have no native vision. | aiskillstore/ | 433 | — | ~1.1k | Automated safety check: Pass | No licence | yesterday |
| 184 | A skill your agent uses when deciding what belongs in a CVPR supplementary upload versus the 8-page body, covering the one-week-later supplement deadline, video and qualitative-result norms in… | brycewang-stanford/ | 1.2k | — | ~1.7k | Automated safety check: Pass | MIT | 13 days ago |
| 185 | A skill your agent uses when deciding whether a computer-vision project should target ECCV — weighing the two-year even-year cadence against CVPR, ICCV, WACV, BMVC, ACCV, NeurIPS, and journal… | brycewang-stanford/ | 1.2k | — | ~1.1k | Automated safety check: Pass | MIT | 13 days ago |
| 186 | A skill your agent uses when deciding whether a computer-vision project should target ICCV, weighing the biennial odd-year cadence and its two-year option cost, what the Marr Prize lineage says… | brycewang-stanford/ | 1.2k | — | ~1.7k | Automated safety check: Pass | MIT | 13 days ago |
| 187 | 覆盖工具 allow/ask/deny、四种 permission mode、审批持久化、Workspace/Hook Trust、敏感文件以及 Hook/Browser 网络边界。 | echoVic/ | 181 | — | ~1.8k | Automated safety check: Pass | MIT | today |
| 188 | 188.AI ML Skills 27 ai & machine learning skills. An agent skill from wentorai/research-plugins. | wentorai/ | 298 | 1 repo | ~993 | Automated safety check: Pass | MIT | 3 mo ago |
| 189 | Stand up a complete, ready-to-run computer-vision analytics stack on Intel hardware with one Docker Compose command — point it at your video sources and an OpenVINO/ONNX model to get live annotated… | open-edge-platform/ | 140 | — | ~4.4k | Automated safety check: Notes | Apache-2.0 | yesterday |
| 190 | Build machine vision inspection systems with MATLAB Visual Inspection Toolbox. | matlab/ | 1.1k | — | ~3.1k | Automated safety check: Pass | Unknown | 2 days ago |
| 191 | 利用多模态AI分析商品主图,提取视觉特征和提示词。当用户提到分析产品图片、从商品图中提取视觉属性、识别产品Listing中的颜色/形状/材质/风格、反推图片提示词、批量视觉特征提取、将产品图信息转化为结构化数据、视觉属性统计、基于图片的商品分类、main image analysis, image feature extraction, visual attribute… | linkfox-ai/ | 107 | — | ~2.2k | Automated safety check: Pass | MIT | 26 days ago |
| 192 | 基于多模态AI的图片识别与分析。当用户想分析、描述、从图片URL中提取信息、image recognition, image analysis, image description, image content understanding, OCR text recognition, visual… | linkfox-ai/ | 107 | — | ~1.7k | Automated safety check: Pass | MIT | 26 days ago |