Search

Python · Computer vision

24 skills found.
Search results
#SkillRepositoryStarsUsed inTokensAuto-checkLicenceUpdated
1

Guide to using Meta's Segment Anything Model for zero-shot image segmentation with point, box or mask prompts, or automatic mask generation.

Orchestra-Research/AI-Research-SKILLs13k9 repos~3.3kAutomated safety check: PassMIT3 mo ago
2

Explains OpenAI's CLIP model for zero-shot image classification, image-text similarity, semantic image search and content moderation, with install steps and code patterns.

Orchestra-Research/AI-Research-SKILLs13k8 repos~1.7kAutomated safety check: PassMIT3 mo ago
3

Composites several moments from real drone footage into one still with ghost trails, then lays out paper figures and an editable PowerPoint file.

XXLiu-HNU/visualize_uav_trajectory242—~535Automated safety check: PassGPL-3.012 days ago
4

Guide to LLaVA for image chat, visual question answering and captioning, with model sizes, CLI and Gradio usage and multi-turn conversation code.

Orchestra-Research/AI-Research-SKILLs13k7 repos~2kAutomated safety check: PassMIT3 mo ago
5

Trains and fine-tunes object detection, image classification and SAM or SAM2 segmentation models on Hugging Face Jobs cloud GPUs and saves the results to the Hub.

huggingface/skills11k1 repo~7.5kAutomated safety check: PassApache-2.07 days ago
6

Interactive click-to-segment using Segment Anything 2 — AI-assisted labeling for Annotation Studio

SharpAI/DeepCamera3.1k—~594Automated safety check: PassMIT22 days ago
7

Call vision models (Doubao, Qwen, DeepSeek, OpenAI) to analyze images.

xiincs/claude-code-vision-skill170—~1.2kAutomated safety check: PassMIT1 mo ago
8

Google Coral Edge TPU — real-time object detection natively (macOS / Linux)

SharpAI/DeepCamera3.1k—~1.2kAutomated safety check: PassMIT22 days ago
9

Low-level SenseNova tools for image generation, image editing, image recognition with a VLM and text optimization with an LLM, meant to be called by higher-level skills rather than directly.

OpenSenseNova/SenseNova-Skills5.7k—~3.2kAutomated safety check: PassMIT21 days ago
10

Google Coral Edge TPU — real-time object detection natively via Windows WSL

SharpAI/DeepCamera3.1k—~1.1kAutomated safety check: PassMIT22 days ago
11

Extracts pixel, video-frame, speech, music and visual-semantic features from image, video and audio files for research datasets, using local tools or the Gemini API.

TyrealQ/q-skills108—~2kAutomated safety check: NotesMIT15 days ago
12

Processes digital pathology whole slide images with histolab: tissue detection, mask creation, tile extraction and dataset preparation for deep learning.

davila7/claude-code-templates32k12 repos~5.1kAutomated safety check: PassMITyesterday
13

Loads pre-trained Hugging Face Transformers models for text, vision and audio tasks, runs inference with pipelines and fine-tunes on custom datasets.

davila7/claude-code-templates32k12 repos~1.2kAutomated safety check: PassMITyesterday
14

Extracts body and hand keypoint trajectories from an authorized reference video, with skeleton previews and confidence data, for pose reference or motion control input.

Pluviobyte/rnskill1.6k—~1kAutomated safety check: PassUnknown17 days ago
15

Review public AlbumentationsX docstrings for useful descriptions, runnable examples, parameter semantics, and related transforms.

albumentations-team/AlbumentationsX567—~534Automated safety check: PassAGPL-3.0yesterday
16

A skill your agent uses when user asks to analyze an image, describe image contents, or answer questions about a picture.

iflytek/iFly-Skills209—~949Automated safety check: PassApache-2.0yesterday
17

Run the full DEFT smart-data-augmentation loop for NVIDIA TAO Grounding DINO object detection: zero-shot baseline inference, KPI analysis, per-class gap analysis, SigLIP embedding of weak images…

NVIDIA/skills3.5k—~3.3kAutomated safety check: NotesApache-2.0yesterday
18

End-to-end ONNX PTQ workflow for AMD Quark — from a .onnx file (and calibration data) to a quantized .onnx output.

amd/Quark181—~4.5kAutomated safety check: PassMIT11 days ago
19
19.Deepstream DevOfficial

NVIDIA DeepStream SDK development with Python pyservicemaker API.

NVIDIA/skills3.5k—~3.3kAutomated safety check: PassApache-2.0yesterday
20

Creates MATLAB interfaces to Python image processing and computer vision models from GitHub repositories or pip-installable packages using MPyReq.

matlab/matlab-agentic-toolkit1.1k—~3.7kAutomated safety check: PassUnknownyesterday
21

Computer vision for bio-image preprocessing, feature detection, real-time microscopy.

jaechang-hits/SciAgent-Skills3701 repo~3.9kAutomated safety check: PassApache-2.010 days ago
22

Guidance for building and training with the Caffe deep learning framework on CIFAR-10 dataset.

lazyFrogLOL/Harness_Engineering128—~1.7kAutomated safety check: PassNo licence4 mo ago
23

Use this sub-skill for Ultralytics YOLO predict workflows, source handling, streaming and batching, Results extraction, saving/plotting/cropping, and thread-safe inference.

VectorSpaceLab/AREX-Skill328—~1.4kAutomated safety check: PassAGPL-3.01 mo ago
24

A skill your agent uses for Ultralytics YOLO package workflows: CLI/Python model usage, data/config setup, train/val, prediction/results, export/deployment, tracking/solutions, model-family…

VectorSpaceLab/AREX-Skill328—~1.2kAutomated safety check: PassAGPL-3.01 mo ago