Search

Computer vision

195 skills found, page 2.
Search results
#SkillRepositoryStarsUsed inTokensAuto-checkLicenceUpdated
49

Understand images and videos with Qwen vision models. An agent skill from QianWen-AI/qianwen-ai.

QianWen-AI/qianwen-ai105—~4.9kAutomated safety check: NotesApache-2.0today
50

Guide for implementing Google Gemini API image understanding - analyze images with captioning, classification, visual QA, object detection, segmentation, and multi-image comparison.

einverne/dotfiles121—~1.6kAutomated safety check: NotesMIT1 mo ago
51

Maintain AlbumentationsX license, CLA, provenance notices, and packaged legal artifacts consistently.

albumentations-team/AlbumentationsX567—~1.3kAutomated safety check: PassAGPL-3.0today
52

Runs NVIDIA TAO Data Services KPI analysis on object detection results, comparing predictions to ground truth and writing per-class precision, recall and AP to a CSV.

NVIDIA/skills3.6k—~2.7kAutomated safety check: NotesApache-2.0yesterday
53
53.Transformers.jsOfficial

Runs pre-trained Hugging Face models in JavaScript or TypeScript with Transformers.js, in browsers or Node.js, Bun and Deno, for text, vision, audio and multimodal tasks.

huggingface/skills11k1 repo~6.2kAutomated safety check: PassApache-2.02 days ago
54

A skill your agent uses when user asks to analyze an image, describe image contents, or answer questions about a picture.

iflytek/iFly-Skills209—~949Automated safety check: PassApache-2.02 days ago
55

Deep expertise in ML/CV model selection, training pipelines, and inference architecture.

alirezarezvani/claude-cto-team117—~3.1kAutomated safety check: PassMIT9 mo ago
56

After completing code changes, runs tests and pre-commit, then iteratively fixes failures until all pass.

albumentations-team/AlbumentationsX567—~847Automated safety check: PassAGPL-3.0today
57

Runs TAO Data Services gap analysis that compares ground-truth and predicted boxes to find weak images by per-class recall, precision and AP50.

NVIDIA/skills3.6k—~1.8kAutomated safety check: NotesApache-2.0yesterday
58

World-class computer vision skill for image/video processing, object detection, segmentation, and visual AI systems.

davila7/claude-code-templates33k2 repos~1.4kAutomated safety check: PassMITtoday
59

Build an end-to-end UAV object detection and telemetry overlay application on Intel hardware using DL Streamer Pipeline Server with MAVLink telemetry.

open-edge-platform/edge-ai-suites140—~2.6kAutomated safety check: NotesApache-2.0today
60

Looks up Microsoft Learn guidance for Azure AI Vision: Image Analysis, Read OCR containers, smart-crop thumbnails, background removal and video frame analysis, plus limits and deployment.

MicrosoftDocs/Agent-Skills776—~1.6kAutomated safety check: PassCC-BY-4.04 days ago
61

Compares a vision-language model's yes/no predictions with ground truth and writes the false-positive and false-negative cases to a JSONL file with a summary report.

NVIDIA/skills3.6k—~1.3kAutomated safety check: NotesApache-2.0yesterday
62

Build image analysis applications with Azure AI Vision SDK for Java.

microsoft/skills3.1k5 repos~2.2kAutomated safety check: PassMITyesterday
63

Azure AI Vision Image Analysis SDK for captions, tags, objects, OCR, people detection, and smart cropping.

microsoft/skills3.1k5 repos~2.5kAutomated safety check: PassMITyesterday
64

Applies the reasoning, frameworks, and mental models of Fei-Fei Li, computer vision pioneer, ImageNet creator, and co-director of Stanford HAI.

K-Dense-AI/mimeo282—~1.8kAutomated safety check: PassMIT1 mo ago
65

Build production computer vision pipelines for object detection, tracking, and video analysis.

curiositech/some_claude_skills244—~4kAutomated safety check: PassMIT1 mo ago
66

Use the repo internal/ directory for anything that must not be committed — scratch files, temporary outputs, local demos, Codex artifacts, or one-off scripts.

albumentations-team/AlbumentationsX567—~338Automated safety check: PassAGPL-3.0today
67

Agent-driven YOLO fine-tuning — annotate, train, export, deploy

SharpAI/DeepCamera3.1k—~985Automated safety check: PassMIT23 days ago
68

Builds with and operates Pi, the minimal terminal coding harness.

K-Dense-AI/scientific-agent-skills48k1 repo~2.1kAutomated safety check: PassMIT5 days ago
69

Turns a parquet of image file paths into a parquet of embeddings with CLIP, SigLIP or a TAO checkpoint, using the TAO Data Services container, ahead of neighbor mining.

NVIDIA/skills3.6k—~2kAutomated safety check: NotesApache-2.0yesterday
70

This skill should be used when user asks to "improve my mAP", "why is my model overfitting", "my training is diverging", "read my results.csv", "interpret my training curves", "my AP50 is good but…

fcakyon/claude-codex-settings1.2k—~1.4kAutomated safety check: PassApache-2.0yesterday
71

Computer vision engineering skill for object detection, image segmentation, and visual AI systems.

alirezarezvani/claude-skills28k1 repo~3.2kAutomated safety check: PassMIT1 mo ago
72

Train or fine-tune vision models on Hugging Face Jobs for detection, classification, and SAM or SAM2 segmentation.

henryalouf/ruflow157—~7.4kAutomated safety check: PassMIT4 mo ago
73
73.Docker Agent RunOfficial

A skill your agent uses when running a Docker Agent with docker agent run, choosing a safety/approval mode, using the --sandbox isolation flag, setting up aliases, or troubleshooting a run (missing…

docker/skills552—~2.2kAutomated safety check: PassApache-2.06 days ago
74

Best practices for image classification tasks. An agent skill from aiming-lab/AutoResearchClaw.

aiming-lab/AutoResearchClaw15k—~304Automated safety check: PassMIT1 mo ago
75

Measure AlbumentationsX runtime changes with paired baseline and candidate benchmarks on the affected routes.

albumentations-team/AlbumentationsX567—~900Automated safety check: PassAGPL-3.0today
76

Use vision models to self-review screenshots against design intent.

dylanfeltus/skills179—~2.4kAutomated safety check: PassMIT22 days ago
77

Run TAO Data Services TMM unique-neighbor matching mining from embedding parquet files for object detection workflows.

NVIDIA/skills3.6k—~2.2kAutomated safety check: NotesApache-2.0yesterday
78

Review an AlbumentationsX transform for correctness, public API coherence, performance, documentation, and test coverage.

albumentations-team/AlbumentationsX567—~973Automated safety check: PassAGPL-3.0today
79

Best practices for object detection tasks. An agent skill from aiming-lab/AutoResearchClaw.

aiming-lab/AutoResearchClaw15k—~257Automated safety check: PassMIT1 mo ago
80

Train object detection, image classification, and SAM or SAM2 segmentation models locally or on Hugging Face Jobs, with dataset validation and results saved to the Hub.

sickn33/agentic-awesome-skills47k1 repo~1.1kAutomated safety check: PassApache-2.0yesterday
81

Build on-device AI features in React Native and Expo apps with React Native ExecuTorch.

software-mansion-labs/skills291—~2.3kAutomated safety check: PassNo licence12 days ago
82

Fine-tune vision-language models (VLMs) with supervised learning on image+text data.

wshobson/agents40k—~2kAutomated safety check: PassMIT5 days ago
83

Implement computer vision features including text recognition (OCR), face detection, barcode scanning, image segmentation, object tracking, and document scanning in iOS apps.

dpearson2699/swift-ios-skills1.2k—~4.7kAutomated safety check: PassUnknown2 mo ago
84

Upgrade a coded website to award-tier, editorially-crafted design using fal.ai.

fal-ai-community/skills251—~1.4kAutomated safety check: PassMIT11 days ago
85

Work with state-of-the-art machine learning models for NLP, computer vision, audio, and multimodal tasks using HuggingFace Transformers.

ynulihao/AgentSkillOS618—~2.9kAutomated safety check: PassNo licence7 mo ago
86

火山视频理解 - 使用火山方舟视频理解 API 分析视频内容。通过 Files API 上传视频(推荐),支持大文件(最大512MB),可用于视频内容分析、物体识别、动作理解等。当用户需要分析视频、理解视频内容、提取视频信息时激活此技能。

freestylefly/canghe-skills461—~1kAutomated safety check: NotesNo licence4 mo ago
87

Run the full DEFT smart-data-augmentation loop for NVIDIA TAO Grounding DINO object detection: zero-shot baseline inference, KPI analysis, per-class gap analysis, SigLIP embedding of weak images…

NVIDIA/skills3.6k—~3.3kAutomated safety check: NotesApache-2.0yesterday
88

Port a published computer vision paper's official code and training recipe onto a customer's own dataset, or diagnose why such a transfer produced bad numbers.

NVIDIA/skills3.6k—~4.2kAutomated safety check: NotesApache-2.0yesterday
89

Grade or filter workflow HDF5 episodes with an OpenAI-compatible vision model.

NVIDIA/skills3.6k1 repo~1.5kAutomated safety check: PassApache-2.0yesterday
90

Lightweight, Chorus-native local specs for Chorus PM workflows in Hermes — a durable local spec .chorus/specs/<slug/spec.md (one per capability/feature) edited in place and NEVER synced (git history…

Chorus-AIDLC/Chorus1.2k—~2.2kAutomated safety check: PassAGPL-3.0yesterday
91

Lightweight, Chorus-native local specs for Chorus PM workflows in Pi — a durable local spec .chorus/specs/<slug/spec.md (one per capability/feature) edited in place and NEVER synced (git history is…

Chorus-AIDLC/Chorus1.2k—~2.1kAutomated safety check: PassAGPL-3.0yesterday
92

Lightweight, Chorus-native local specs for Chorus PM workflows on OpenClaw — a durable local spec .chorus/specs/<slug/spec.md (one per capability/feature) edited in place and NEVER synced (git…

Chorus-AIDLC/Chorus1.2k—~2.4kAutomated safety check: PassAGPL-3.0yesterday
93

Lightweight, Chorus-native local specs for Chorus PM workflows — a durable local spec .chorus/specs/<slug/spec.md (one per capability/feature) edited in place and NEVER synced (git history is its…

Chorus-AIDLC/Chorus1.2k—~2.1kAutomated safety check: PassAGPL-3.0yesterday
94

Lightweight, Chorus-native local specs for Chorus PM workflows on dsh — a durable local spec .chorus/specs/<slug/spec.md (one per capability/feature) edited in place and NEVER synced (git history is…

Chorus-AIDLC/Chorus1.2k—~2.4kAutomated safety check: PassAGPL-3.0yesterday
95

A skill your agent uses to bring a supported object-detection vision model from HuggingFace or NVIDIA NGC into an NVIDIA DeepStream pipeline with end-to-end automation: ONNX download, SafeTensors…

NVIDIA/skills3.6k—~3.6kAutomated safety check: PassApache-2.0yesterday
96

Sparse4D for multi-camera temporal 3D object detection and tracking.

NVIDIA/skills3.6k—~3.8kAutomated safety check: NotesApache-2.0yesterday