Agent skill

Describe Image

by Wide-Moat in Wide-Moat/open-computer-use

Describe images (charts, diagrams, tables, screenshots) using Vision AI.

Custom licenceAuto-check passedData & Analytics

Install Describe Image

skills CLI
$ npx skills add Wide-Moat/open-computer-use --skill describe-image -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install Wide-Moat/open-computer-use describe-image --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/Wide-Moat/open-computer-use.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/public/describe-image .claude/skills/describe-image && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
describe-image
GitHub stars
126
Token cost
~732 tokens
SKILL.md length
147 words
Files
2 (incl. scripts)
Skills in repo
5
Repo updated
First seen
Licence
Custom licence

At a glance

Describe images (charts, diagrams, tables, screenshots) using Vision AI.

  • Tasks that involve Diagrams
  • SKILL.md covers Quick Start, Parameters, Supported Image Types and Examples, plus 2 more sections
  • Runs Python scripts from its folder; calls python; needs VISION_API_KEY
  • Tasks that involve Data pipelines and ETL

What it does

Describe Image is an agent skill from Wide-Moat/open-computer-use. Describe images (charts, diagrams, tables, screenshots) using Vision AI. Use as fallback when you cannot read an image file directly. Supports batch processing of folders.

Its SKILL.md is about 730 tokens, which your agent loads only when the skill is triggered. The skill folder holds 2 other files, including scripts (for example `scripts/describe.py`).

It sits in Data & Analytics, covering Diagrams and Data pipelines and ETL. The repository describes itself as: ARCHIVED — superseded by Wide Moat.

When your agent uses it

  • Tasks that involve Diagrams
  • Tasks that involve Data pipelines and ETL

Example prompts

  • “/describe-image”

Requirements

  • Python 3
  • A credential in VISION_API_KEY

What it can do on your machine

Read from SKILL.md and the folder at commit 1ddf832. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Ships 1 file in scripts/ (Python), which the agent can run.

    Shell commands in SKILL.md call:

    • python

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names these keys or tokens, usually read from environment variables:

    • VISION_API_KEY

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Describe Image loads about 732 tokens when it runs. Until then it costs about 47 tokens; SKILL.md has 147 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~47
When it runs · the whole SKILL.md, loaded when a task matches
~732

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); the scripts in this folder are not scanned.

SKILL.md

Its licence (Custom licence) doesn't allow us to republish the file, so here is its outline and opening line. It has 147 words (~732 tokens).

“Analyze images using Vision AI and generate detailed text descriptions in Russian.”

— opening of SKILL.md by Wide-Moat, Custom licence
name
describe-image

Read the full SKILL.md on GitHub

Files

SKILL.md and 1 other file (scripts) in skills/public/describe-image of Wide-Moat/open-computer-use.

  • SKILL.md
  • scripts/describe.py

Open the folder on GitHubat commit 1ddf832

Compare with similar skills

Describe Image next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Describe Image compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Describe Image this skillWide-Moat/open-computer-use126—~732Automated safety check: PassCustom licence
Dynamic ArchifyOWWZO/ai-agent1891 repos~5.5kAutomated safety check: PassMIT
Technical Diagram Skillmingchen666/Reviva244—~5.7kAutomated safety check: PassMIT
Erd Studio Setupliam-machine/erd-studio165—~8.6kAutomated safety check: PassCustom licence
Archifymolvqingtai/WebChat2.6k—~5.6kAutomated safety check: PassMIT
Diagrammerdavila7/claude-code-templates33k—~494Automated safety check: PassMIT

Similar skills

  • Dynamic Archify

    OWWZO/ai-agent

    Create professional architecture, workflow, sequence, data-flow, and lifecycle/state diagrams as standalone animated HTML files with SVG graphics, flowing animation effects, a built-in dark/light…

    189 GitHub starsUsed in 1 repo~5.5k tokens
    Data & AnalyticsAuto-check passed
  • Technical Diagram Skill

    mingchen666/Reviva

    Create professional technical diagrams as standalone animated HTML files with SVG graphics, flowing effects, dark/light theme toggle, and export support.

    244 GitHub stars~5.7k tokensUpdated today
    Data & AnalyticsAuto-check passed
  • Erd Studio Setup

    liam-machine/erd-studio

    Friendly, step-by-step setup for ERD Studio in an existing dbt project, for people who may be new to dbt or data modelling.

    165 GitHub stars~8.6k tokensUpdated today
    Data & AnalyticsAuto-check passed
  • Archify

    molvqingtai/WebChat

    Create professional architecture, workflow, sequence, data-flow, and lifecycle/state diagrams as standalone HTML files with SVG graphics, a built-in dark/light theme toggle, and one-click export to…

    2.6k GitHub stars~5.6k tokensUpdated 27 days ago
    DevelopmentAuto-check passed
  • Diagrammer

    davila7/claude-code-templates

    Render clean blueprint-style SVG diagrams from JSON specs. An agent skill from davila7/claude-code-templates.

    33k GitHub stars~494 tokensUpdated today
    DevelopmentAuto-check passed
  • Archify Diagram Builder

    Unclecheng-li/AI_Animation

    Builds validated architecture, workflow, sequence, data-flow and lifecycle diagrams as standalone interactive HTML from a small JSON spec, with optional motion and image export.

    1.5k GitHub starsUsed in 2 repos~4.1k tokens
    DevelopmentAuto-check passed

More from Wide-Moat/open-computer-use

  • GitLab Explorer

    Wide-Moat/open-computer-use

    Explore GitLab repositories using glab CLI and git commands.

    126 GitHub stars~891 tokensUpdated 4 days ago
    Auto-check passed
  • Sub Agent

    Wide-Moat/open-computer-use

    COSTLY: Spawns a separate sub-agent CLI session. An agent skill from Wide-Moat/open-computer-use.

    126 GitHub stars~1.9k tokensUpdated 4 days ago
    Auto-check passed
  • File Reading

    Wide-Moat/open-computer-use

    A skill your agent uses when a file has been uploaded but its content is NOT in your context — only its path at /mnt/user-data/uploads/ is listed in an uploadedfiles block.

    126 GitHub starsUsed in 1 repo~3.1k tokens
    Auto-check passed
  • PDF Reading

    Wide-Moat/open-computer-use

    A skill your agent uses when you need to read, inspect, or extract content from PDF files — especially when file content is NOT in your context and you need to read it from disk.

    126 GitHub starsUsed in 1 repo~2.7k tokens
    Auto-check passed

Questions about Describe Image

What does Describe Image do?

Describe images (charts, diagrams, tables, screenshots) using Vision AI. Describe Image is an agent skill from Wide-Moat/open-computer-use. Describe images (charts, diagrams, tables, screenshots) using Vision AI.

When should I use Describe Image?

Describe Image fits situations like: tasks that involve Diagrams; tasks that involve Data pipelines and ETL.

How do I install Describe Image in Claude Code?

Run `npx skills add Wide-Moat/open-computer-use --skill describe-image -a claude-code`. Or copy the skill folder (skills/public/describe-image in Wide-Moat/open-computer-use) into .claude/skills/describe-image in your project. Claude Code loads it when a task matches its description.

How do I install Describe Image in Codex?

Run `npx skills add Wide-Moat/open-computer-use --skill describe-image -a codex`. Or copy the skill folder (skills/public/describe-image in Wide-Moat/open-computer-use) into .agents/skills/describe-image in your project. Codex loads it when a task matches its description.

Can I use Describe Image in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add Wide-Moat/open-computer-use --skill describe-image -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/describe-image, .gemini/skills/describe-image, .github/skills/describe-image and .opencode/skills/describe-image in your project.

What does Describe Image need to run?

Going by SKILL.md and its folder, Describe Image needs Python for the scripts in its folder, the command-line tools its instructions call (python) and credentials named VISION_API_KEY. Our summary lists: Python 3; A credential in VISION_API_KEY.

Does Describe Image access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is Describe Image safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. The check reads SKILL.md only: the scripts in the folder are not scanned, so read them before running anything.

What licence does Describe Image use?

Describe Image has a licence file (the repository's licence) that doesn't match a standard licence. Read it on GitHub before reusing the skill.

How many tokens does Describe Image use?

About 732 tokens (SKILL.md is roughly 2.9k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Describe Image?

Skills that share tags, products or a category with Describe Image: Dynamic Archify (OWWZO/ai-agent, 189 stars), Technical Diagram Skill (mingchen666/Reviva, 244 stars), Erd Studio Setup (liam-machine/erd-studio, 165 stars) and Archify (molvqingtai/WebChat, 2.6k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Describe Image?

Wide-Moat (a GitHub organization) maintains it in Wide-Moat/open-computer-use, which has 126 GitHub stars. The repository holds 5 skills in this directory. The repository was last updated on October 5, 2026.

Source: Wide-Moat/open-computer-use on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.