Azure AI Vision Imageanalysis Py
microsoft/skills
Azure AI Vision Image Analysis SDK for captions, tags, objects, OCR, people detection, and smart cropping.
Looks up Microsoft Learn guidance for Azure AI Vision: Image Analysis, Read OCR containers, smart-crop thumbnails, background removal and video frame analysis, plus limits and deployment.
$ npx skills add MicrosoftDocs/Agent-Skills --skill azure-ai-vision -a claude-codeProject install by default; add -g for ~/.claude/skills/.
$ gh skill install MicrosoftDocs/Agent-Skills azure-ai-vision --agent claude-codeProject scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).
$ git clone --depth 1 https://github.com/MicrosoftDocs/Agent-Skills.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/azure-ai-vision .claude/skills/azure-ai-vision && rm -rf skills-srcUse ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.
Claude Code skills documentation · loads skills from .claude/skills/
Install the "azure-ai-vision" agent skill from https://github.com/MicrosoftDocs/Agent-Skills/tree/main/skills/azure-ai-vision into .claude/skills/azure-ai-vision/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "azure-ai-vision", then confirm the skill loads.Claude Code copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$skill-installer install https://github.com/MicrosoftDocs/Agent-Skills/tree/main/skills/azure-ai-visionType this inside Codex. $skill-installer <name> installs a curated skill from openai/skills. The installer writes to $CODEX_HOME/skills (default ~/.codex/skills). Restart Codex if the skill does not show up.
$ npx skills add MicrosoftDocs/Agent-Skills --skill azure-ai-vision -a codexProject install goes to .agents/skills/; add -g for ~/.codex/skills/.
$ gh skill install MicrosoftDocs/Agent-Skills azure-ai-vision --agent codexProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/MicrosoftDocs/Agent-Skills.git skills-src && mkdir -p .agents/skills && cp -r skills-src/skills/azure-ai-vision .agents/skills/azure-ai-vision && rm -rf skills-srcUse ~/.agents/skills/ instead of .agents/skills for a personal install.
Codex skills documentation · loads skills from .agents/skills/
Install the "azure-ai-vision" agent skill from https://github.com/MicrosoftDocs/Agent-Skills/tree/main/skills/azure-ai-vision into .agents/skills/azure-ai-vision/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "azure-ai-vision", then confirm the skill loads.Codex copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add MicrosoftDocs/Agent-Skills --skill azure-ai-vision -a cursorProject install goes to .agents/skills/; add -g for ~/.cursor/skills/.
$ gh skill install MicrosoftDocs/Agent-Skills azure-ai-vision --agent cursorProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/MicrosoftDocs/Agent-Skills.git skills-src && mkdir -p .cursor/skills && cp -r skills-src/skills/azure-ai-vision .cursor/skills/azure-ai-vision && rm -rf skills-srcUse ~/.cursor/skills/ instead of .cursor/skills for a personal install.
Cursor skills documentation · loads skills from .cursor/skills/, .agents/skills/, .claude/skills/, .codex/skills/
Install the "azure-ai-vision" agent skill from https://github.com/MicrosoftDocs/Agent-Skills/tree/main/skills/azure-ai-vision into .cursor/skills/azure-ai-vision/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "azure-ai-vision", then confirm the skill loads.Cursor copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gemini skills install https://github.com/MicrosoftDocs/Agent-Skills.git --path skills/azure-ai-vision--scope user (default) or --scope workspace; --path is the subfolder of the repo that holds the skill; --consent skips the security confirmation prompt.
$ npx skills add MicrosoftDocs/Agent-Skills --skill azure-ai-vision -a gemini-cliProject install goes to .agents/skills/; add -g for ~/.gemini/skills/.
$ gh skill install MicrosoftDocs/Agent-Skills azure-ai-vision --agent gemini-cliProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/MicrosoftDocs/Agent-Skills.git skills-src && mkdir -p .gemini/skills && cp -r skills-src/skills/azure-ai-vision .gemini/skills/azure-ai-vision && rm -rf skills-srcUse ~/.gemini/skills/ instead of .gemini/skills for a personal install, then run /skills reload.
Gemini CLI skills documentation · loads skills from .gemini/skills/, .agents/skills/
Install the "azure-ai-vision" agent skill from https://github.com/MicrosoftDocs/Agent-Skills/tree/main/skills/azure-ai-vision into .gemini/skills/azure-ai-vision/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "azure-ai-vision", then confirm the skill loads.Gemini CLI copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gh skill install MicrosoftDocs/Agent-Skills azure-ai-visionInstalls for Copilot at project scope by default; add --scope user for a personal install. Preview a skill first with gh skill preview. Needs GitHub CLI 2.90.0 or later (public preview).
$ npx skills add MicrosoftDocs/Agent-Skills --skill azure-ai-vision -a github-copilotProject install goes to .agents/skills/; add -g for ~/.copilot/skills/.
$ git clone --depth 1 https://github.com/MicrosoftDocs/Agent-Skills.git skills-src && mkdir -p .github/skills && cp -r skills-src/skills/azure-ai-vision .github/skills/azure-ai-vision && rm -rf skills-srcUse ~/.copilot/skills/ instead of .github/skills for a personal install. Commit .github/skills so cloud agent and code review can use it.
GitHub Copilot skills documentation · loads skills from .github/skills/, .claude/skills/, .agents/skills/
Install the "azure-ai-vision" agent skill from https://github.com/MicrosoftDocs/Agent-Skills/tree/main/skills/azure-ai-vision into .github/skills/azure-ai-vision/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "azure-ai-vision", then confirm the skill loads.GitHub Copilot copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add MicrosoftDocs/Agent-Skills --skill azure-ai-vision -a opencodeOpenCode documents no install command of its own. Project install goes to .agents/skills/; add -g for ~/.config/opencode/skills/.
$ gh skill install MicrosoftDocs/Agent-Skills azure-ai-vision --agent opencodeProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/MicrosoftDocs/Agent-Skills.git skills-src && mkdir -p .opencode/skills && cp -r skills-src/skills/azure-ai-vision .opencode/skills/azure-ai-vision && rm -rf skills-srcUse ~/.config/opencode/skills/ instead of .opencode/skills for a personal install.
OpenCode skills documentation · loads skills from .opencode/skills/, .claude/skills/, .agents/skills/
Install the "azure-ai-vision" agent skill from https://github.com/MicrosoftDocs/Agent-Skills/tree/main/skills/azure-ai-vision into .opencode/skills/azure-ai-vision/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "azure-ai-vision", then confirm the skill loads.OpenCode copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
azure-ai-visionLooks up Microsoft Learn guidance for Azure AI Vision: Image Analysis, Read OCR containers, smart-crop thumbnails, background removal and video frame analysis, plus limits and deployment.
The skill is an index into Azure AI Vision documentation rather than a set of rules. A category index points to line ranges or reference files for five areas: decision making (migrating Image Analysis and Read OCR apps, including the move from v2.x to v3.x APIs), limits and quotas, configuration of Read OCR containers and Azure Blob Storage access, integration and coding patterns (OCR, embeddings, thumbnails, background removal, domain models, live video frame analysis) and deployment of the Read OCR container locally or on-premises.
It fetches current documentation pages over the network, preferably with the Microsoft Docs MCP fetch tool and otherwise with a web page fetch, and tells the agent to suggest an update when the skill's generated date is more than three months old. It is scoped to Azure AI Vision and points elsewhere for Custom Vision, Video Indexer, Document Intelligence and Immersive Reader.
Read from SKILL.md and the folder at commit ba74e8f. It shows what the files ask for, not the result of running them.
Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.
From allowed-tools in the SKILL.md frontmatter.
No scripts in the folder and no shell commands in SKILL.md.
From the folder's file list and the shell code blocks in SKILL.md.
Links to these hosts (documentation or services it may open):
learn.microsoft.comgithub.comFrom URLs in SKILL.md, links to its own repository left out.
Names no API keys, tokens, secrets or passwords.
From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.
Requires network access. Uses mcp_microsoftdocs:microsoft_docs_fetch or fetch_webpage to retrieve documentation.
From compatibility in the SKILL.md frontmatter.
Azure AI Vision Reference loads about 1.6k tokens when it runs. Until then it costs about 144 tokens; SKILL.md has 451 words of instructions outside code blocks.
Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.
The automated check found no risky patterns in SKILL.md.
Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.
The full file from MicrosoftDocs/Agent-Skills at commit ba74e8f, republished under its CC-BY-4.0 licence (© MicrosoftDocs). 451 words, ~1,594 tokens.
.claude/skills/azure-ai-vision/SKILL.md (or your agent's skills folder).This skill provides expert guidance for Azure AI Vision. Covers decision making, limits & quotas, configuration, integrations & coding patterns, and deployment. It combines local quick-reference content with remote documentation fetching capabilities.
IMPORTANT for Agent: Use the Category Index below to locate relevant sections. For categories with line ranges (e.g.,
L35-L120), useread_filewith the specified lines. For categories with file links (e.g.,[security.md](security.md)), useread_fileon the linked reference file
IMPORTANT for Agent: If
metadata.generated_atis more than 3 months old, suggest the user pull the latest version from the repository. Ifmcp_microsoftdocstools are not available, suggest the user install it: Installation Guide
This skill requires network access to fetch documentation content:
mcp_microsoftdocs:microsoft_docs_fetch with query string from=learn-agent-skill. Returns Markdown.fetch_webpage with query string from=learn-agent-skill&accept=text/markdown. Returns Markdown.| Category | Lines | Description |
|---|---|---|
| Decision Making | L33-L39 | Guidance on migrating and upgrading Azure Vision Image Analysis and Read OCR apps/containers, including choosing migration paths and moving from v2.x to v3.x APIs. |
| Limits & Quotas | L40-L50 | Limits, thresholds, and taxonomies for Image Analysis: category lists, adult content scores, object/people detection constraints, smart-crop behavior, and OCR language support. |
| Configuration | L51-L56 | Configuring Vision Read OCR containers and setting up Azure Blob Storage access for image input, including environment settings, storage permissions, and connection details. |
| Integrations & Coding Patterns | L57-L67 | How to call and configure Azure Vision/Read APIs and SDKs for OCR, embeddings, thumbnails, background removal, domain models, and live video frame analysis. |
| Deployment | L68-L71 | Installing, configuring, and running the Azure AI Vision Read OCR container locally or on-premises, including prerequisites, deployment steps, and runtime settings. |
| Topic | URL |
|---|---|
| Choose migration path from Azure Vision Image Analysis | https://learn.microsoft.com/en-us/azure/ai-services/computer-vision/migration-options |
| Migrate to Azure Vision Read OCR container v3.x | https://learn.microsoft.com/en-us/azure/ai-services/computer-vision/read-container-migration-guide |
| Upgrade applications from Read v2.x to v3.0 | https://learn.microsoft.com/en-us/azure/ai-services/computer-vision/upgrade-api-versions |
| Topic | URL |
|---|---|
| Configure Azure Vision Read OCR containers | https://learn.microsoft.com/en-us/azure/ai-services/computer-vision/computer-vision-resource-container-config |
| Configure Azure Blob Storage for Vision image retrieval | https://learn.microsoft.com/en-us/azure/ai-services/computer-vision/how-to/blob-storage-search |
| Topic | URL |
|---|---|
| Call domain-specific models with Azure Vision | https://learn.microsoft.com/en-us/azure/ai-services/computer-vision/concept-detecting-domain-content |
| Analyze live video frames with Azure Vision API | https://learn.microsoft.com/en-us/azure/ai-services/computer-vision/how-to/analyze-video |
| Call and configure Image Analysis 3.2 API | https://learn.microsoft.com/en-us/azure/ai-services/computer-vision/how-to/call-analyze-image |
| Call and configure Image Analysis 4.0 API | https://learn.microsoft.com/en-us/azure/ai-services/computer-vision/how-to/call-analyze-image-40 |
| Call and configure Azure Vision Read v3.2 API | https://learn.microsoft.com/en-us/azure/ai-services/computer-vision/how-to/call-read-api |
| Use multimodal embeddings for image retrieval | https://learn.microsoft.com/en-us/azure/ai-services/computer-vision/how-to/image-retrieval |
| Use OCR client libraries for text extraction | https://learn.microsoft.com/en-us/azure/ai-services/computer-vision/quickstarts-sdk/client-library |
| Topic | URL |
|---|---|
| Install and run Azure Vision Read OCR container | https://learn.microsoft.com/en-us/azure/ai-services/computer-vision/computer-vision-how-to-install-containers |
© MicrosoftDocs, CC-BY-4.0. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file
Just SKILL.md in skills/azure-ai-vision of MicrosoftDocs/Agent-Skills.
Open the folder on GitHubat commit ba74e8f
Azure AI Vision Reference next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.
| Skill | Stars | Used in | Tokens | Auto-check | Licence | Repo updated |
|---|---|---|---|---|---|---|
| Azure AI Vision Reference this skillMicrosoftDocs/Agent-Skills | 777 | — | ~1.6k | Automated safety check: Pass | CC-BY-4.0 | |
| Azure AI Vision Imageanalysis Pymicrosoft/skills | 3.1k | 6 repos | ~2.5k | Automated safety check: Pass | MIT | |
| Microsoft Foundrymicrosoft/GitHub-Copilot-for-Azure | 255 | 1 repos | ~6.7k | Automated safety check: Pass | MIT | |
| Ak Inityaalalabs/agent-kernel | 191 | — | ~2.8k | Automated safety check: Pass | Apache-2.0 | |
| UltralyticsVectorSpaceLab/AREX-Skill | 328 | — | ~1.2k | Automated safety check: Pass | AGPL-3.0 | |
| Cvatmajiayu000/claude-skill-registry | 666 | 1 repos | ~909 | Automated safety check: Pass | MIT |
microsoft/skills
Azure AI Vision Image Analysis SDK for captions, tags, objects, OCR, people detection, and smart cropping.
microsoft/GitHub-Copilot-for-Azure
Build, deploy, evaluate, optimize, fine-tune, and manage Microsoft Foundry agents, models, and resources end to end.
yaalalabs/agent-kernel
Scaffold a new Agent Kernel project from scratch. An agent skill from yaalalabs/agent-kernel.
VectorSpaceLab/AREX-Skill
A skill your agent uses for Ultralytics YOLO package workflows: CLI/Python model usage, data/config setup, train/val, prediction/results, export/deployment, tracking/solutions, model-family…
majiayu000/claude-skill-registry
Operate CVAT for computer-vision annotation, dataset workflows, SDK/CLI automation, auto-annotation, and self-hosted deployment.
microsoft/skills
Build image analysis applications with Azure AI Vision SDK for Java.
MicrosoftDocs/Agent-Skills
Expert knowledge for Azure AI Personalizer development including troubleshooting, decision making, security, configuration, and integrations & coding patterns.
MicrosoftDocs/Agent-Skills
Guides Azure solution design by category, from reference architectures and design patterns to technology choices and migrations, fetching current Microsoft Learn pages over the network.
MicrosoftDocs/Agent-Skills
Reference guidance for Azure Advisor work: recommendations, alerts and digests, workbooks, RBAC access and sovereign-cloud limits, fetched from Microsoft Learn.
MicrosoftDocs/Agent-Skills
Expert knowledge for Azure Analysis Services development including troubleshooting.
MicrosoftDocs/Agent-Skills
Expert knowledge for Azure AI Anomaly Detector development including troubleshooting, best practices, limits & quotas, configuration, and deployment.
MicrosoftDocs/Agent-Skills
Expert knowledge for Azure Anyscale On Azure development including limits & quotas, security, configuration, and deployment.
Categories
Looks up Microsoft Learn guidance for Azure AI Vision: Image Analysis, Read OCR containers, smart-crop thumbnails, background removal and video frame analysis, plus limits and deployment. The skill is an index into Azure AI Vision documentation rather than a set of rules.x APIs), limits and quotas, configuration of Read OCR containers and Azure Blob Storage access, integration and coding patterns (OCR, embeddings, thumbnails, background removal, domain models, live video frame analysis) and deployment of the Read OCR container locally or on-premises.
Azure AI Vision Reference fits situations like: migrating an Image Analysis or Read OCR app to a newer API version; checking Azure AI Vision limits, quotas and supported OCR languages; running the Read OCR container locally or on-premises; calling Image Analysis for thumbnails or background removal.
Run `npx skills add MicrosoftDocs/Agent-Skills --skill azure-ai-vision -a claude-code`. Or copy the skill folder (skills/azure-ai-vision in MicrosoftDocs/Agent-Skills) into .claude/skills/azure-ai-vision in your project. Claude Code loads it when a task matches its description.
Run `npx skills add MicrosoftDocs/Agent-Skills --skill azure-ai-vision -a codex`. Or copy the skill folder (skills/azure-ai-vision in MicrosoftDocs/Agent-Skills) into .agents/skills/azure-ai-vision in your project. Codex loads it when a task matches its description.
Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add MicrosoftDocs/Agent-Skills --skill azure-ai-vision -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/azure-ai-vision, .gemini/skills/azure-ai-vision, .github/skills/azure-ai-vision and .opencode/skills/azure-ai-vision in your project.
SKILL.md names no scripts, command-line tools or credentials: Azure AI Vision Reference is instructions for the agent only. Our summary lists: Network access to Microsoft Learn documentation; The Microsoft Docs MCP server, or a web page fetch tool. Compatibility (from SKILL.md): Requires network access. Uses mcp_microsoftdocs:microsoft_docs_fetch or fetch_webpage to retrieve documentation..
SKILL.md names 2 domains. As links in the text: learn.microsoft.com and github.com. This is read from the text; nothing was executed.
Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.
Azure AI Vision Reference is published under the CC-BY-4.0 licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.
About 1.6k tokens (SKILL.md is roughly 6.4k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.
Skills that share tags, products or a category with Azure AI Vision Reference: Azure AI Vision Imageanalysis Py (microsoft/skills, 3.1k stars), Microsoft Foundry (microsoft/GitHub-Copilot-for-Azure, 255 stars), Ak Init (yaalalabs/agent-kernel, 191 stars) and Ultralytics (VectorSpaceLab/AREX-Skill, 328 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.
MicrosoftDocs (a GitHub organization, an official publisher) maintains it in MicrosoftDocs/Agent-Skills, which has 777 GitHub stars. The repository holds 149 skills in this directory. The repository was last updated on October 5, 2026.
Source: MicrosoftDocs/Agent-Skills on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.