Topic · AI & LLM Engineering

Best computer vision skills, page 5

Skills #193–206 of 206, ranked by score.

Computer vision skills, ranked

Ranked by score. Sort bymost stars,trending,newest,recently updated

Computer vision skills, ranked
#SkillRepositoryStarsUsed inTokensAuto-checkLicenceUpdated
193

利用多模态AI分析商品主图,提取视觉特征和提示词。当用户提到分析产品图片、从商品图中提取视觉属性、识别产品Listing中的颜色/形状/材质/风格、反推图片提示词、批量视觉特征提取、将产品图信息转化为结构化数据、视觉属性统计、基于图片的商品分类、main image analysis, image feature extraction, visual attribute…

linkfox-ai/linkfox-skills106—~2.2kAutomated safety check: PassMIT23 days ago
194

基于多模态AI的图片识别与分析。当用户想分析、描述、从图片URL中提取信息、image recognition, image analysis, image description, image content understanding, OCR text recognition, visual…

linkfox-ai/linkfox-skills106—~1.7kAutomated safety check: PassMIT23 days ago
195

Expert knowledge for Azure AI Custom Vision development including best practices, decision making, limits & quotas, security, integrations & coding patterns, and deployment.

MicrosoftDocs/Agent-Skills775—~1.6kAutomated safety check: PassCC-BY-4.02 days ago
196

Generate article or newsletter thumbnail candidates using the Gemini API from inside Claude Code.

mohitagw15856/pm-claude-skills1.4k—~6.3kAutomated safety check: PassMITyesterday
197

Unlock deep video understanding with qwen-ai on ClawHub. An agent skill from LeoYeAI/openclaw-master-skills.

LeoYeAI/openclaw-master-skills2.2k—~4.2kAutomated safety check: PassMIT2 mo ago
198

Process many issues hands-off in a row: resolve a queue, then run each through the yolo chain under one up-front confirm.

jepegit/cellpy109—~3.5kAutomated safety check: PassMITyesterday
199

Review open GitHub issues and apply labels (extendable kinds; v1: yolo).

jepegit/cellpy109—~1.7kAutomated safety check: PassMITyesterday
200

Chain capture → plan → build → close yolo for a small, low-risk issue under one consolidated confirm.

jepegit/cellpy109—~1.6kAutomated safety check: PassMITyesterday
201

Elite Computer Vision Engineer skill with expertise in deep learning for images and video (CNNs, Transformers), object detection (YOLO, DETR), segmentation, OCR, and production CV deployment…

theneoai/awesome-skills183—~2.2kAutomated safety check: PassMIT4 mo ago
202

Expert-level Perception Algorithm Engineer with deep knowledge of 3D object detection (PointPillars, VoxelNet, BEVFusion, DETR3D), semantic segmentation (BEV), multi-camera fusion (BEVFormer), LiDAR…

theneoai/awesome-skills183—~3.4kAutomated safety check: PassUnknown4 mo ago
203

Expert robot perception engineer specializing in 3D point cloud processing, multi-modal sensor fusion (camera+LiDAR+IMU), real-time SLAM, and edge-optimized deep learning inference via TensorRT/ONNX…

theneoai/awesome-skills183—~1.5kAutomated safety check: PassMIT4 mo ago
204

Computer vision implementation covering image classification, object detection (YOLO), image segmentation, OCR (Tesseract), face detection, image preprocessing, data augmentation, transfer learning…

FerroxLabs/wayland608—~3.9kAutomated safety check: PassApache-2.0yesterday
205

Hands-on computer vision development covering image classification with transfer learning, object detection with YOLO and Faster R-CNN, semantic and instance segmentation, OpenCV image processing…

FerroxLabs/wayland608—~4.6kAutomated safety check: PassApache-2.0yesterday
206

Multimodal AI pipeline design covering vision-language models, text-audio integration, image generation, model selection across modalities, preprocessing pipelines, fusion architectures, and…

FerroxLabs/wayland608—~3.2kAutomated safety check: PassApache-2.0yesterday