Agent skill

Senior Computer Vision

by davila7 in davila7/claude-code-templates

World-class computer vision skill for image/video processing, object detection, segmentation, and visual AI systems.

MITAuto-check passedAI & LLM Engineering

Install Senior Computer Vision

skills CLI
$ npx skills add davila7/claude-code-templates --skill senior-computer-vision -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install davila7/claude-code-templates senior-computer-vision --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/davila7/claude-code-templates.git skills-src && mkdir -p .claude/skills && cp -r skills-src/cli-tool/components/skills/development/senior-computer-vision .claude/skills/senior-computer-vision && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
senior-computer-vision
GitHub stars
32k
Used in
2 other repos
Token cost
~1.4k tokens
SKILL.md length
440 words
Files
7 (incl. scripts, references)
Skills in repo
478
Repo updated
First seen
Licence
MIT

At a glance

World-class computer vision skill for image/video processing, object detection, segmentation, and visual AI systems.

  • Works in 3 steps: Computer Vision Architectures → Object Detection Optimization → Production Vision Systems
  • Building vision AI systems
  • SKILL.md covers Quick Start, Core Expertise, Tech Stack and Reference Documentation, plus 7 more sections
  • Runs Python scripts from its folder; calls python, kubectl and docker

What it does

Senior Computer Vision is an agent skill from davila7/claude-code-templates. World-class computer vision skill for image/video processing, object detection, segmentation, and visual AI systems. Expertise in PyTorch, OpenCV, YOLO, SAM, diffusion models, and vision transformers. Includes 3D vision, video analysis, real-time processing, and production deployment. Use when building vision AI systems, implementing object detection, training custom vision models, or optimizing inference pipelines.

Its SKILL.md is about 1.4k tokens, which your agent loads only when the skill is triggered. The skill folder holds 8 other files, including scripts and reference files (for example `references/computer_vision_architectures.md`, `references/object_detection_optimization.md` and `references/production_vision_systems.md`).

It sits in AI & LLM Engineering, covering Computer vision. It works with PyTorch and OpenCV. The repository describes itself as: CLI tool for configuring and monitoring Claude Code. The licence is MIT.

When your agent uses it

  • Building vision AI systems
  • Implementing object detection
  • Training custom vision models
  • Optimizing inference pipelines

Example prompts

  • “/senior-computer-vision”

Requirements

  • Python 3
  • Docker

Workflow steps

3 steps, taken from the step headings in SKILL.md.

  1. Computer Vision Architectures
  2. Object Detection Optimization
  3. Production Vision Systems

What it can do on your machine

Read from SKILL.md and the folder at commit 46b4d8b. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Ships 3 files in scripts/ (Python), which the agent can run.

    Shell commands in SKILL.md call:

    • python
    • kubectl
    • docker
    • helm

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md. Its commands use kubectl, docker and helm, which can reach the network depending on how they are called.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Senior Computer Vision loads about 1.4k tokens when it runs, and up to ~2.5k if it reads all its reference files. Until then it costs about 111 tokens; SKILL.md has 440 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~111
When it runs · the whole SKILL.md, loaded when a task matches
~1.4k
With references · SKILL.md plus every file in references/, read only if the agent opens them
~2.5k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); the scripts in this folder are not scanned.

SKILL.md

The full file from davila7/claude-code-templates at commit 46b4d8b, republished under its MIT licence (© davila7). 440 words, ~1,399 tokens.

Download SKILL.mdSave it as .claude/skills/senior-computer-vision/SKILL.md (or your agent's skills folder). This skill also uses 6 other files; get the full folder from GitHub.
name
senior-computer-vision
description
World-class computer vision skill for image/video processing, object detection, segmentation, and visual AI systems. Expertise in PyTorch, OpenCV, YOLO, SAM, diffusion models, and vision transformers. Includes 3D vision, video analysis, real-time processing, and production deployment. Use when building vision AI systems, implementing object detection, training custom vision models, or optimizing inference pipelines.

Senior Computer Vision Engineer

World-class senior computer vision engineer skill for production-grade AI/ML/Data systems.

Quick Start

Main Capabilities
bash
# Core Tool 1
python scripts/vision_model_trainer.py --input data/ --output results/

# Core Tool 2  
python scripts/inference_optimizer.py --target project/ --analyze

# Core Tool 3
python scripts/dataset_pipeline_builder.py --config config.yaml --deploy

Core Expertise

This skill covers world-class capabilities in:

  • Advanced production patterns and architectures
  • Scalable system design and implementation
  • Performance optimization at scale
  • MLOps and DataOps best practices
  • Real-time processing and inference
  • Distributed computing frameworks
  • Model deployment and monitoring
  • Security and compliance
  • Cost optimization
  • Team leadership and mentoring

Tech Stack

Languages: Python, SQL, R, Scala, Go ML Frameworks: PyTorch, TensorFlow, Scikit-learn, XGBoost Data Tools: Spark, Airflow, dbt, Kafka, Databricks LLM Frameworks: LangChain, LlamaIndex, DSPy Deployment: Docker, Kubernetes, AWS/GCP/Azure Monitoring: MLflow, Weights & Biases, Prometheus Databases: PostgreSQL, BigQuery, Snowflake, Pinecone

Reference Documentation

1. Computer Vision Architectures

Comprehensive guide available in references/computer_vision_architectures.md covering:

  • Advanced patterns and best practices
  • Production implementation strategies
  • Performance optimization techniques
  • Scalability considerations
  • Security and compliance
  • Real-world case studies
2. Object Detection Optimization

Complete workflow documentation in references/object_detection_optimization.md including:

  • Step-by-step processes
  • Architecture design patterns
  • Tool integration guides
  • Performance tuning strategies
  • Troubleshooting procedures
3. Production Vision Systems

Technical reference guide in references/production_vision_systems.md with:

  • System design principles
  • Implementation examples
  • Configuration best practices
  • Deployment strategies
  • Monitoring and observability

Production Patterns

Pattern 1: Scalable Data Processing

Enterprise-scale data processing with distributed computing:

  • Horizontal scaling architecture
  • Fault-tolerant design
  • Real-time and batch processing
  • Data quality validation
  • Performance monitoring
Pattern 2: ML Model Deployment

Production ML system with high availability:

  • Model serving with low latency
  • A/B testing infrastructure
  • Feature store integration
  • Model monitoring and drift detection
  • Automated retraining pipelines
Pattern 3: Real-Time Inference

High-throughput inference system:

  • Batching and caching strategies
  • Load balancing
  • Auto-scaling
  • Latency optimization
  • Cost optimization

Best Practices

Show full SKILL.md (181 more words)Show less
Development
  • Test-driven development
  • Code reviews and pair programming
  • Documentation as code
  • Version control everything
  • Continuous integration
Production
  • Monitor everything critical
  • Automate deployments
  • Feature flags for releases
  • Canary deployments
  • Comprehensive logging
Team Leadership
  • Mentor junior engineers
  • Drive technical decisions
  • Establish coding standards
  • Foster learning culture
  • Cross-functional collaboration

Performance Targets

Latency:

  • P50: < 50ms
  • P95: < 100ms
  • P99: < 200ms

Throughput:

  • Requests/second: > 1000
  • Concurrent users: > 10,000

Availability:

  • Uptime: 99.9%
  • Error rate: < 0.1%

Security & Compliance

  • Authentication & authorization
  • Data encryption (at rest & in transit)
  • PII handling and anonymization
  • GDPR/CCPA compliance
  • Regular security audits
  • Vulnerability management

Common Commands

bash
# Development
python -m pytest tests/ -v --cov
python -m black src/
python -m pylint src/

# Training
python scripts/train.py --config prod.yaml
python scripts/evaluate.py --model best.pth

# Deployment
docker build -t service:v1 .
kubectl apply -f k8s/
helm upgrade service ./charts/

# Monitoring
kubectl logs -f deployment/service
python scripts/health_check.py

Resources

  • Advanced Patterns: references/computer_vision_architectures.md
  • Implementation Guide: references/object_detection_optimization.md
  • Technical Reference: references/production_vision_systems.md
  • Automation Scripts: scripts/ directory

Senior-Level Responsibilities

As a world-class senior professional:

  1. Technical Leadership

    • Drive architectural decisions
    • Mentor team members
    • Establish best practices
    • Ensure code quality
  2. Strategic Thinking

    • Align with business goals
    • Evaluate trade-offs
    • Plan for scale
    • Manage technical debt
  3. Collaboration

    • Work across teams
    • Communicate effectively
    • Build consensus
    • Share knowledge
  4. Innovation

    • Stay current with research
    • Experiment with new approaches
    • Contribute to community
    • Drive continuous improvement
  5. Production Excellence

    • Ensure high availability
    • Monitor proactively
    • Optimize performance
    • Respond to incidents

© davila7, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

SKILL.md and 6 other files (scripts, references) in cli-tool/components/skills/development/senior-computer-vision of davila7/claude-code-templates.

  • SKILL.md
  • references/computer_vision_architectures.md
  • references/object_detection_optimization.md
  • references/production_vision_systems.md
  • scripts/dataset_pipeline_builder.py
  • scripts/inference_optimizer.py
  • scripts/vision_model_trainer.py

Open the folder on GitHubat commit 46b4d8b

Used in 2 other repositories

We found 3 copies of this SKILL.md (exact, near-identical or edited) in other folders, from 2 other GitHub owners. This page covers the copy in davila7/claude-code-templates, which our catalogue first saw on October 7, 2026.

Compare with similar skills

Senior Computer Vision next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Senior Computer Vision compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Senior Computer Vision this skilldavila7/claude-code-templates32k2 repos~1.4kAutomated safety check: PassMIT
Torch Performance Optimizationalbumentations-team/albucore123—~895Automated safety check: PassMIT
Computer Vision Pipelinecuriositech/some_claude_skills243—~4kAutomated safety check: PassMIT
Senior Computer Visionalirezarezvani/claude-skills28k1 repos~3.2kAutomated safety check: PassMIT
Tao Finetune ClipNVIDIA/skills3.5k—~4kAutomated safety check: NotesApache-2.0
Tao Finetune Huggingface ModelNVIDIA/skills3.5k—~4.9kAutomated safety check: NotesApache-2.0

Similar skills

  • Torch Performance Optimization

    albumentations-team/albucore

    Optimize or review eager CPU-only Albucore PyTorch runtime paths with benchmark-backed decisions.

    123 GitHub stars~895 tokensUpdated 4 days ago
    AI & LLM EngineeringAuto-check passed
  • Computer Vision Pipeline

    curiositech/some_claude_skills

    Build production computer vision pipelines for object detection, tracking, and video analysis.

    243 GitHub stars~4k tokensUpdated 1 mo ago
    AI & LLM EngineeringAuto-check passed
  • Senior Computer Vision

    alirezarezvani/claude-skills

    Computer vision engineering skill for object detection, image segmentation, and visual AI systems.

    28k GitHub starsUsed in 1 repo~3.2k tokens
    AI & LLM EngineeringAuto-check passed
  • Tao Finetune Clip

    NVIDIA/skills

    Official

    CLIP vision-language model for image-text retrieval, zero-shot classification, embedding extraction, ONNX export, and TensorRT deployment.

    3.5k GitHub stars~4k tokensUpdated yesterday
    AI & LLM EngineeringAuto-check: notes
  • Official

    Fine-tune any HuggingFace CV / VLM / LLM model on local NVIDIA GPUs inside an NGC PyTorch container when no dedicated TAO model skill matches.

    3.5k GitHub stars~4.9k tokensUpdated yesterday
    AI & LLM EngineeringAuto-check: notes
  • Official

    PyTorch-based TAO image classification. An agent skill from NVIDIA/skills.

    3.5k GitHub stars~3.6k tokensUpdated yesterday
    AI & LLM EngineeringAuto-check: notes

More from davila7/claude-code-templates

All 478 skills in this repo
  • Perplexity Web Search

    davila7/claude-code-templates

    Runs web-grounded searches through Perplexity's Sonar models over OpenRouter for current events, recent literature and cited facts beyond the model's training cutoff.

    32k GitHub starsUsed in 11 repos~3.5k tokens
    Auto-check: notes
  • Neuropixels Data Analysis

    davila7/claude-code-templates

    Analyzes Neuropixels recordings from SpikeGLX or Open Ephys through preprocessing, drift correction, Kilosort4 spike sorting, quality metrics and curation.

    32k GitHub starsUsed in 9 repos~2.8k tokens
    Auto-check passed
  • Scientific Venue Templates

    davila7/claude-code-templates

    Supplies LaTeX templates and formatting rules for journals, conferences, posters, and grant proposals, then can check a draft against them.

    32k GitHub starsUsed in 9 repos~5.1k tokens
    Auto-check: notes
  • Brand Voice Content Creator

    davila7/claude-code-templates

    Analyzes a brand's existing writing to lock in a consistent voice, then builds SEO blog posts and platform-specific social content around it.

    32k GitHub starsUsed in 3 repos~1.9k tokens
    Auto-check passed
  • CAPA Officer

    davila7/claude-code-templates

    Guides corrective and preventive action (CAPA) work in a quality management system, from initiation and root cause analysis through effectiveness verification.

    32k GitHub starsUsed in 1 repo~2k tokens
    Auto-check passed
  • Fda Consultant Specialist

    davila7/claude-code-templates

    Senior FDA consultant and specialist for medical device companies including HIPAA compliance and requirement management.

    32k GitHub starsUsed in 1 repo~2.7k tokens
    Auto-check passed

Works with

Questions about Senior Computer Vision

What does Senior Computer Vision do?

World-class computer vision skill for image/video processing, object detection, segmentation, and visual AI systems. Senior Computer Vision is an agent skill from davila7/claude-code-templates. World-class computer vision skill for image/video processing, object detection, segmentation, and visual AI systems.

When should I use Senior Computer Vision?

Senior Computer Vision fits situations like: building vision AI systems; implementing object detection; training custom vision models; optimizing inference pipelines.

How do I install Senior Computer Vision in Claude Code?

Run `npx skills add davila7/claude-code-templates --skill senior-computer-vision -a claude-code`. Or copy the skill folder (cli-tool/components/skills/development/senior-computer-vision in davila7/claude-code-templates) into .claude/skills/senior-computer-vision in your project. Claude Code loads it when a task matches its description.

How do I install Senior Computer Vision in Codex?

Run `npx skills add davila7/claude-code-templates --skill senior-computer-vision -a codex`. Or copy the skill folder (cli-tool/components/skills/development/senior-computer-vision in davila7/claude-code-templates) into .agents/skills/senior-computer-vision in your project. Codex loads it when a task matches its description.

Can I use Senior Computer Vision in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add davila7/claude-code-templates --skill senior-computer-vision -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/senior-computer-vision, .gemini/skills/senior-computer-vision, .github/skills/senior-computer-vision and .opencode/skills/senior-computer-vision in your project.

What does Senior Computer Vision need to run?

Going by SKILL.md and its folder, Senior Computer Vision needs Python for the scripts in its folder and the command-line tools its instructions call (python, kubectl, docker and helm). Our summary lists: Python 3; Docker.

Does Senior Computer Vision access the network?

SKILL.md contains no URLs. Its commands use docker, which can reach the network depending on how they are called. This is read from the text; nothing was executed.

Is Senior Computer Vision safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. The check reads SKILL.md only: the scripts in the folder are not scanned, so read them before running anything.

What licence does Senior Computer Vision use?

Senior Computer Vision is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Senior Computer Vision use?

About 1.4k tokens (SKILL.md is roughly 5.6k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full. Its references folder adds about 1.1k tokens, read only when the agent opens those files.

What are the alternatives to Senior Computer Vision?

Skills that share tags, products or a category with Senior Computer Vision: Torch Performance Optimization (albumentations-team/albucore, 123 stars), Computer Vision Pipeline (curiositech/some_claude_skills, 243 stars), Senior Computer Vision (alirezarezvani/claude-skills, 28k stars) and Tao Finetune Clip (NVIDIA/skills, 3.5k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Senior Computer Vision?

davila7 (a GitHub user) maintains it in davila7/claude-code-templates, which has 32,483 GitHub stars. The repository holds 478 skills in this directory. The repository was last updated on October 9, 2026.

Source: davila7/claude-code-templates on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.