Topic · AI & LLM Engineering
Best computer vision skills, page 5
Computer vision skills, ranked
Ranked by score. Sort bymost stars,trending,newest,recently updated
| # | Skill | Repository | Stars | Used in | Tokens | Auto-check | Licence | Updated |
|---|---|---|---|---|---|---|---|---|
| 193 | 利用多模态AI分析商品主图,提取视觉特征和提示词。当用户提到分析产品图片、从商品图中提取视觉属性、识别产品Listing中的颜色/形状/材质/风格、反推图片提示词、批量视觉特征提取、将产品图信息转化为结构化数据、视觉属性统计、基于图片的商品分类、main image analysis, image feature extraction, visual attribute… | linkfox-ai/ | 106 | — | ~2.2k | Automated safety check: Pass | MIT | 23 days ago |
| 194 | 基于多模态AI的图片识别与分析。当用户想分析、描述、从图片URL中提取信息、image recognition, image analysis, image description, image content understanding, OCR text recognition, visual… | linkfox-ai/ | 106 | — | ~1.7k | Automated safety check: Pass | MIT | 23 days ago |
| 195 | Expert knowledge for Azure AI Custom Vision development including best practices, decision making, limits & quotas, security, integrations & coding patterns, and deployment. | MicrosoftDocs/ | 775 | — | ~1.6k | Automated safety check: Pass | CC-BY-4.0 | 2 days ago |
| 196 | Generate article or newsletter thumbnail candidates using the Gemini API from inside Claude Code. | mohitagw15856/ | 1.4k | — | ~6.3k | Automated safety check: Pass | MIT | yesterday |
| 197 | 197.Qwen AI Unlock deep video understanding with qwen-ai on ClawHub. An agent skill from LeoYeAI/openclaw-master-skills. | LeoYeAI/ | 2.2k | — | ~4.2k | Automated safety check: Pass | MIT | 2 mo ago |
| 198 | 198.Iflow Cycle Process many issues hands-off in a row: resolve a queue, then run each through the yolo chain under one up-front confirm. | jepegit/ | 109 | — | ~3.5k | Automated safety check: Pass | MIT | yesterday |
| 199 | 199.Iflow Review Review open GitHub issues and apply labels (extendable kinds; v1: yolo). | jepegit/ | 109 | — | ~1.7k | Automated safety check: Pass | MIT | yesterday |
| 200 | 200.Iflow Yolo Chain capture → plan → build → close yolo for a small, low-risk issue under one consolidated confirm. | jepegit/ | 109 | — | ~1.6k | Automated safety check: Pass | MIT | yesterday |
| 201 | Elite Computer Vision Engineer skill with expertise in deep learning for images and video (CNNs, Transformers), object detection (YOLO, DETR), segmentation, OCR, and production CV deployment… | theneoai/ | 183 | — | ~2.2k | Automated safety check: Pass | MIT | 4 mo ago |
| 202 | Expert-level Perception Algorithm Engineer with deep knowledge of 3D object detection (PointPillars, VoxelNet, BEVFusion, DETR3D), semantic segmentation (BEV), multi-camera fusion (BEVFormer), LiDAR… | theneoai/ | 183 | — | ~3.4k | Automated safety check: Pass | Unknown | 4 mo ago |
| 203 | Expert robot perception engineer specializing in 3D point cloud processing, multi-modal sensor fusion (camera+LiDAR+IMU), real-time SLAM, and edge-optimized deep learning inference via TensorRT/ONNX… | theneoai/ | 183 | — | ~1.5k | Automated safety check: Pass | MIT | 4 mo ago |
| 204 | 204.Computer Vision Computer vision implementation covering image classification, object detection (YOLO), image segmentation, OCR (Tesseract), face detection, image preprocessing, data augmentation, transfer learning… | FerroxLabs/ | 608 | — | ~3.9k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 205 | Hands-on computer vision development covering image classification with transfer learning, object detection with YOLO and Faster R-CNN, semantic and instance segmentation, OpenCV image processing… | FerroxLabs/ | 608 | — | ~4.6k | Automated safety check: Pass | Apache-2.0 | yesterday |
| 206 | Multimodal AI pipeline design covering vision-language models, text-audio integration, image generation, model selection across modalities, preprocessing pipelines, fusion architectures, and… | FerroxLabs/ | 608 | — | ~3.2k | Automated safety check: Pass | Apache-2.0 | yesterday |
Explore related skills
More topics in AI & LLM Engineering
- Building AI agents525
- Deep learning408
- Embeddings381
- LLM inference and serving364
- Retrieval-augmented generation360
- Prompt engineering350
- Fine-tuning313
- LLM evaluation303
- Speech recognition and synthesis272
- Structured output and tool calling271
- LLM cost and token optimization256
- LLM API integration218
- LLM observability217
- Model routing and gateways217
- LLM guardrails208
- Model hubs and datasets180
- GPU and accelerator computing171
- Diffusion and image models167
- Natural language processing141
- Reinforcement learning67
- AI interpretability23