Topic · AI & LLM Engineering
Best computer vision skills, page 4
Computer vision skills, ranked
Ranked by score. Sort bymost stars,trending,newest,recently updated
| # | Skill | Repository | Stars | Used in | Tokens | Auto-check | Licence | Updated |
|---|---|---|---|---|---|---|---|---|
| 145 | Advanced CV for infrastructure inspection including forest fire detection, wildfire precondition assessment, roof inspection, hail damage analysis, thermal imaging, and 3D Gaussian Splatting… | curiositech/ | 243 | 1 repo | ~2.2k | Automated safety check: Pass | MIT | 1 mo ago |
| 146 | Integrate a HuggingFace Computer Vision model into the NVIDIA TAO Toolkit ecosystem (tao-core config, tao-pytorch trainer, tao-deploy TensorRT pipeline). | NVIDIA/ | 3.5k | — | ~4.5k | Automated safety check: Notes | Apache-2.0 | yesterday |
| 147 | Based on computer vision, analyzes pet health indicators such as feeding frequency, drinking frequency, excretion status, mental state, vomiting behavior, and limping abnormalities through… | LeoYeAI/ | 2.2k | — | ~1.6k | Automated safety check: Pass | MIT | 2 mo ago |
| 148 | Computer vision engineering for object detection, segmentation, and visual AI, covering CNN and Vision Transformer architectures and ONNX/TensorRT deployment. | borghei/ | 881 | — | ~1.8k | Automated safety check: Pass | MIT | yesterday |
| 149 | Display images and annotations for image processing, computer vision, and visual inspection. | matlab/ | 1.1k | — | ~3.8k | Automated safety check: Pass | Unknown | yesterday |
| 150 | Display 3-D image volumes, medical image volumes, surface meshes, and annotations for 3-D image processing. | matlab/ | 1.1k | — | ~4.8k | Automated safety check: Pass | Unknown | yesterday |
| 151 | Register 3-D point clouds using ICP, NDT, LOAM, FGR, phase correlation, and CPD algorithms. | matlab/ | 1.1k | — | ~4.4k | Automated safety check: Pass | Unknown | yesterday |
| 152 | 152.Object Counter Count occurrences of an object in the image using computer vision algorithm. | benchflow-ai/ | 1.8k | 1 repo | ~140 | Automated safety check: Pass | Apache-2.0 | 2 mo ago |
| 153 | 153.Scholar Compute Design and execute computational social science analyses across 11 modules: text-as-data/NLP (STM, BERTopic, Wordfish, BERT, conText embedding regression, LLM annotation + DSL bias correction… | joshzyj/ | 168 | — | ~15k | Automated safety check: Pass | Unknown | 20 days ago |
| 154 | Computer vision for bio-image preprocessing, feature detection, real-time microscopy. | jaechang-hits/ | 370 | 1 repo | ~3.9k | Automated safety check: Pass | Apache-2.0 | 9 days ago |
| 155 | Deploy and maintain image understanding (OCR + local VLM + cloud VL) and image generation for Codex connected to text-only models like DeepSeek. | aiskillstore/ | 430 | — | ~814 | Automated safety check: Notes | MIT | yesterday |
| 156 | Research, write, code, analyze, present, and develop proposals for thermal-fluid mechanical engineering work with source-aware rigor. | ccplugins/ | 968 | — | ~2.9k | Automated safety check: Pass | Apache-2.0 | 1 mo ago |
| 157 | 157.Gemini CLI Integrate Gemini AI CLI into Claude Code for AI collaboration, code analysis, and tool execution. | majiayu000/ | 666 | 1 repo | ~4.1k | Automated safety check: Notes | MIT | yesterday |
| 158 | 158.Gemini Tools Execute and manage Gemini CLI built-in tools including file operations, web search, shell commands, and memory. | majiayu000/ | 666 | 1 repo | ~3k | Automated safety check: Notes | MIT | yesterday |
| 159 | 159.Ops Business operations command center. An agent skill from davepoon/buildwithclaude. | davepoon/ | 3.6k | — | ~999 | Automated safety check: Notes | MIT | 2 days ago |
| 160 | 160.Ops Yolo YOLO mode. An agent skill from davepoon/buildwithclaude. | davepoon/ | 3.6k | — | ~2.9k | Automated safety check: Notes | MIT | 2 days ago |
| 161 | 161.Ffmpeg A skill your agent uses for local FFmpeg/FFprobe media inspection, remuxing, transcoding, filtering, evidence-bounded video review, transcript-assisted editorial plans, edit decision lists… | magnus919/ | 113 | — | ~3k | Automated safety check: Pass | MIT | 2 days ago |
| 162 | A skill your agent uses when image reading, screenshot analysis, visual diff, design mockup analysis, pasted image handling, or Z.AI Vision MCP fails, times out, returns 429/Too Many Requests, or… | skuramatata/ | 114 | — | ~1k | Automated safety check: Pass | No licence | 3 mo ago |
| 163 | A skill your agent uses when targeting Asian Conference on Computer Vision (ACCV) or deciding whether a computer-science manuscript fits this venue. | franklee16/ | 223 | 1 repo | ~1.9k | Automated safety check: Pass | No licence | 20 days ago |
| 164 | A skill your agent uses when targeting British Machine Vision Conference (BMVC) or deciding whether a computer-science manuscript fits this venue. | franklee16/ | 223 | 1 repo | ~1.9k | Automated safety check: Pass | No licence | 20 days ago |
| 165 | A skill your agent uses when targeting IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) or deciding whether a computer-science manuscript fits this venue. | franklee16/ | 223 | 1 repo | ~2k | Automated safety check: Pass | No licence | 20 days ago |
| 166 | A skill your agent uses when targeting European Conference on Computer Vision (ECCV) or deciding whether a computer-science manuscript fits this venue. | franklee16/ | 223 | 1 repo | ~1.9k | Automated safety check: Pass | No licence | 20 days ago |
| 167 | A skill your agent uses when targeting IEEE Transactions on Pattern Analysis and Machine Intelligence (TPAMI) or deciding whether a computer vision or machine learning manuscript fits this archival… | franklee16/ | 223 | 1 repo | ~2k | Automated safety check: Pass | No licence | 20 days ago |
| 168 | A skill your agent uses when targeting IEEE/CVF International Conference on Computer Vision (ICCV) or deciding whether a computer-science manuscript fits this venue. | franklee16/ | 223 | 1 repo | ~1.9k | Automated safety check: Pass | No licence | 20 days ago |
| 169 | A skill your agent uses when targeting IEEE/CVF Winter Conference on Applications of Computer Vision (WACV) or deciding whether a computer-science manuscript fits this venue. | franklee16/ | 223 | 1 repo | ~2k | Automated safety check: Pass | No licence | 20 days ago |
| 170 | 170.Caffe Cifar 10 Guidance for building and training with the Caffe deep learning framework on CIFAR-10 dataset. | lazyFrogLOL/ | 128 | — | ~1.7k | Automated safety check: Pass | No licence | 4 mo ago |
| 171 | A skill your agent uses whenever the user wants to find, shortlist, vet, or enrich US AI/ML/data consulting firms (consultancies) — AI/ML development, MLOps, generative AI / LLM apps (RAG, chatbots… | jeremylongshore/ | 2.8k | — | ~4k | Automated safety check: Notes | MIT | today |
| 172 | Apply computer vision research methods, models, and evaluation tools | wentorai/ | 298 | 1 repo | ~1.6k | Automated safety check: Pass | MIT | 3 mo ago |
| 173 | Guide to Transformer architectures for NLP and computer vision | wentorai/ | 298 | 1 repo | ~2.1k | Automated safety check: Pass | MIT | 3 mo ago |
| 174 | Build OCR pipelines in MATLAB using the ocr() function. An agent skill from matlab/matlab-agentic-toolkit. | matlab/ | 1.1k | — | ~5.2k | Automated safety check: Pass | Unknown | yesterday |
| 175 | Analyze videos with Google Gemini API (summaries, Q&A, transcription with timestamps + visual context, scene/timeline detection, video clipping, FPS control, multi-video comparison, and YouTube URL… | benchflow-ai/ | 1.8k | — | ~2.4k | Automated safety check: Pass | Apache-2.0 | 2 mo ago |
| 176 | 176.Openai Vision Analyze images and multi-frame sequences using OpenAI GPT vision models | benchflow-ai/ | 1.8k | — | ~5k | Automated safety check: Pass | Apache-2.0 | 2 mo ago |
| 177 | 177.Detection A skill your agent uses for PaddleViT object detection workflows with DETR, Swin, or PVTv2: validate COCO data, select configs, build/train/evaluate models, reason about transforms, losses… | VectorSpaceLab/ | 328 | — | ~2.4k | Automated safety check: Pass | Apache-2.0 | 1 mo ago |
| 178 | 178.Imgaug A skill your agent uses when working with imgaug image augmentation pipelines, aligned annotations, stochastic parameters, dtype/data utilities, or multicore augmentation for computer-vision data. | VectorSpaceLab/ | 328 | — | ~1.2k | Automated safety check: Pass | MIT | 1 mo ago |
| 179 | Use this sub-skill for Ultralytics YOLO predict workflows, source handling, streaming and batching, Results extraction, saving/plotting/cropping, and thread-safe inference. | VectorSpaceLab/ | 328 | — | ~1.4k | Automated safety check: Pass | AGPL-3.0 | 1 mo ago |
| 180 | 180.Mmdetection A skill your agent uses when working with MMDetection 3.x for object detection, instance/panoptic segmentation, tracking-adjacent configs, model zoo configs, inference, visualization… | VectorSpaceLab/ | 328 | — | ~1.2k | Automated safety check: Pass | Apache-2.0 | 1 mo ago |
| 181 | 181.Swin Transformer Use this repo skill for Microsoft Swin-Transformer image-classification model, config, data, checkpoint, SimMIM, Swin-MoE, and optional CUDA acceleration workflows. | VectorSpaceLab/ | 328 | — | ~1.2k | Automated safety check: Pass | MIT | 1 mo ago |
| 182 | A skill your agent uses for Ultralytics track mode, tracker YAML selection, persistent object IDs, and YOLO Solutions such as counting, heatmaps, speed, queue, region, similarity search, and… | VectorSpaceLab/ | 328 | — | ~871 | Automated safety check: Pass | AGPL-3.0 | 1 mo ago |
| 183 | 183.Ultralytics A skill your agent uses for Ultralytics YOLO package workflows: CLI/Python model usage, data/config setup, train/val, prediction/results, export/deployment, tracking/solutions, model-family… | VectorSpaceLab/ | 328 | — | ~1.2k | Automated safety check: Pass | AGPL-3.0 | 1 mo ago |
| 184 | Guide safe video-file, webcam, and optional half-precision demo use for pytorch-yolo-v3. | VectorSpaceLab/ | 328 | — | ~626 | Automated safety check: Pass | Apache-2.0 | 1 mo ago |
| 185 | 185.Yolov5 A skill your agent uses for clone-run Ultralytics YOLOv5 workflows: detection, segmentation, classification, export, benchmarks, datasets, weights, and Flask REST serving. | VectorSpaceLab/ | 328 | — | ~1.6k | Automated safety check: Pass | AGPL-3.0 | 1 mo ago |
| 186 | Paddle-based deep learning workflows from the course materials, including DNN/RNN text-style baselines and the CNN/LeNet image classification case using folder-labeled digit images. | Drchronx/ | 135 | — | ~516 | Automated safety check: Pass | Unknown | 4 mo ago |
| 187 | 187.Pp Twelvelabs Printing Press CLI for Twelvelabs. An agent skill from mvanhorn/printing-press-library. | mvanhorn/ | 2.1k | — | ~3.7k | Automated safety check: Notes | Apache-2.0 | yesterday |
| 188 | 188.Mindspeed Mm Vlm Universal VLM (vision-language understanding model) training guide for Huawei Ascend NPU using MindSpeed-MM. | ascend-ai-coding/ | 174 | — | ~5.2k | Automated safety check: Pass | No licence | today |
| 189 | A skill your agent uses when deciding what belongs in a CVPR supplementary upload versus the 8-page body, covering the one-week-later supplement deadline, video and qualitative-result norms in… | brycewang-stanford/ | 1.2k | — | ~1.7k | Automated safety check: Pass | MIT | 11 days ago |
| 190 | A skill your agent uses when deciding whether a computer-vision project should target ECCV — weighing the two-year even-year cadence against CVPR, ICCV, WACV, BMVC, ACCV, NeurIPS, and journal… | brycewang-stanford/ | 1.2k | — | ~1.1k | Automated safety check: Pass | MIT | 11 days ago |
| 191 | A skill your agent uses when deciding whether a computer-vision project should target ICCV, weighing the biennial odd-year cadence and its two-year option cost, what the Marr Prize lineage says… | brycewang-stanford/ | 1.2k | — | ~1.7k | Automated safety check: Pass | MIT | 11 days ago |
| 192 | 覆盖工具 allow/ask/deny、四种 permission mode、审批持久化、Workspace/Hook Trust、敏感文件以及 Hook/Browser 网络边界。 | echoVic/ | 181 | — | ~1.8k | Automated safety check: Pass | MIT | 11 days ago |
Explore related skills
Category
More topics in AI & LLM Engineering
- Building AI agents563
- Deep learning415
- Embeddings386
- LLM inference and serving372
- Prompt engineering360
- Retrieval-augmented generation358
- Fine-tuning309
- LLM evaluation308
- Speech recognition and synthesis308
- Structured output and tool calling276
- LLM cost and token optimization259
- LLM API integration255
- Model routing and gateways255
- LLM observability240
- LLM guardrails221
- Model hubs and datasets180
- GPU and accelerator computing176
- Diffusion and image models166
- Natural language processing131
- Reinforcement learning66
- AI interpretability23