Agent skill

Image Analysis

by countbot-ai in countbot-ai/CountBot

图片分析与识别,可分析本地图片、网络图片、视频、文件。适用于 OCR、物体识别、场景理解等。当用户发送图片或要求分析图片时必须使用此技能。

MITAuto-check passed

Install Image Analysis

skills CLI
$ npx skills add countbot-ai/CountBot --skill image-analysis -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install countbot-ai/CountBot image-analysis --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/countbot-ai/CountBot.git skills-src && mkdir -p .claude/skills && cp -r skills-src/workspace/skills/image-analysis .claude/skills/image-analysis && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
image-analysis
GitHub stars
781
Used in
1 other repo
Token cost
~532 tokens
SKILL.md length
43 words
Files
5 (incl. scripts)
Skills in repo
11
Repo updated
First seen
Licence
MIT

At a glance

图片分析与识别,可分析本地图片、网络图片、视频、文件。适用于 OCR、物体识别、场景理解等。当用户发送图片或要求分析图片时必须使用此技能。

  • SKILL.md covers 配置, 命令行调用, AI 调用场景 and 模型选择, plus 1 more section
  • Runs Python scripts from its folder; calls python3

What it does

Image Analysis is an agent skill from countbot-ai/CountBot. 图片分析与识别,可分析本地图片、网络图片、视频、文件。适用于 OCR、物体识别、场景理解等。当用户发送图片或要求分析图片时必须使用此技能。

Its SKILL.md is about 530 tokens, which your agent loads only when the skill is triggered. The skill folder holds 5 other files, including scripts (for example `scripts/config.json`, `scripts/vision.py` and `scripts/vision_manager.py`).

The repository describes itself as: 更适配中文用户的轻量开源AI Agent | 国产大模型Coding plan支持 | 兼容OpenClaw Skills生态| 已接入微信ClawBot/微博龙虾/飞书/钉钉/QQ/小智AI/Telegram/deepseek-v4。 The licence is MIT.

Example prompts

  • “/image-analysis”

Requirements

  • Python 3

What it can do on your machine

Read from SKILL.md and the folder at commit 3c26f11. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Ships 4 files in scripts/ (Python), which the agent can run.

    Shell commands in SKILL.md call:

    • python3

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    Links to these hosts (documentation or services it may open):

    • open.bigmodel.cn
    • help.aliyun.com

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Image Analysis loads about 532 tokens when it runs. Until then it costs about 21 tokens; SKILL.md has 43 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~21
When it runs · the whole SKILL.md, loaded when a task matches
~532

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); the scripts in this folder are not scanned.

SKILL.md

The full file from countbot-ai/CountBot at commit 3c26f11, republished under its MIT licence (© countbot-ai). 43 words, ~532 tokens.

Download SKILL.mdSave it as .claude/skills/image-analysis/SKILL.md (or your agent's skills folder). This skill also uses 4 other files; get the full folder from GitHub.
name
image-analysis
description
图片分析与识别,可分析本地图片、网络图片、视频、文件。适用于 OCR、物体识别、场景理解等。当用户发送图片或要求分析图片时必须使用此技能。
homepage
https://github.com/countbot-ai/CountBot

图片分析与识别

支持智谱 GLM-4V 和千问 Qwen-VL 两种视觉模型。

当用户发送图片或要求分析图片时,必须使用此技能,不要使用 PIL、pytesseract 等其他方法。

配置

编辑 skills/image-analysis/scripts/config.json:

json
{
  "default_model": "zhipu",
  "zhipu": {
    "api_key": "your-zhipu-api-key",
    "model": "glm-4.6v-flash"
  },
  "qwen": {
    "api_key": "your-qwen-api-key",
    "model": "qwen3-vl-plus"
  }
}

API Key 获取:

命令行调用

bash
# 分析本地图片(最常用)
python3 skills/image-analysis/scripts/vision.py analyze --image 图片路径 --prompt "描述图片内容"

# 分析网络图片
python3 skills/image-analysis/scripts/vision.py analyze --image https://example.com/image.jpg --prompt "描述图片"

# 多图对比
python3 skills/image-analysis/scripts/vision.py analyze --image img1.jpg --image img2.jpg --prompt "对比差异"

# 指定模型
python3 skills/image-analysis/scripts/vision.py analyze --image image.jpg --prompt "描述图片" --model qwen

# 开启思考模式(仅智谱,提升准确度)
python3 skills/image-analysis/scripts/vision.py analyze --image image.jpg --prompt "详细分析" --thinking

# 视频分析
python3 skills/image-analysis/scripts/vision.py analyze --video video.mp4 --prompt "总结视频内容"

# JSON 输出
python3 skills/image-analysis/scripts/vision.py analyze --image image.jpg --prompt "描述图片" --json

AI 调用场景

用户发送图片后,系统下载到本地(如 data/temp/images/xxx.jpg):

bash
# 图片描述
python3 skills/image-analysis/scripts/vision.py analyze --image data/temp/images/xxx.jpg --prompt "描述这张图片的内容"

# OCR 识别
python3 skills/image-analysis/scripts/vision.py analyze --image data/temp/images/xxx.jpg --prompt "提取图片中的所有文字信息"

# 物体定位(开启思考模式)
python3 skills/image-analysis/scripts/vision.py analyze --image data/temp/images/xxx.jpg --prompt "找出物体位置,返回坐标" --thinking

模型选择

场景推荐
简单描述任意
复杂推理、物体定位智谱 + --thinking
高精度识别、文档解析千问
成本敏感智谱(免费)

注意事项

  • 本地图片自动转 Base64,支持 jpg/png/gif/webp/bmp
  • 智谱图片限制 5MB,像素不超过 6000x6000
  • 千问不支持同时处理图片、视频和文件
  • 思考模式会增加响应时间但提升准确度

© countbot-ai, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

SKILL.md and 4 other files (scripts) in workspace/skills/image-analysis of countbot-ai/CountBot.

  • SKILL.md
  • scripts/config.json
  • scripts/config.json.example
  • scripts/vision.py
  • scripts/vision_manager.py

Open the folder on GitHubat commit 3c26f11

Used in 1 other repository

We found 2 copies of this SKILL.md (exact, near-identical or edited) in other folders, from 1 other GitHub owner. This page covers the copy in countbot-ai/CountBot, which our catalogue first saw on October 7, 2026.

Compare with similar skills

Image Analysis next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Image Analysis compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Image Analysis this skillcountbot-ai/CountBot7811 repos~532Automated safety check: PassMIT
OCR Web Service AutomationComposioHQ/awesome-claude-skills77k3 repos~760Automated safety check: PassNone
Yao OCR ToolsYaoApp/yao8.1k—~1.6kAutomated safety check: PassCustom licence
Doc OCRdavepoon/buildwithclaude3.6k—~244Automated safety check: PassMIT
Image OCRaipoch/medical-research-skills2k—~1.7kAutomated safety check: PassMIT
Post OCR Cleanupbrycewang-stanford/Auto-Empirical-Research-Skills4.5k—~4.7kAutomated safety check: PassCustom licence

Similar skills

  • OCR Web Service Automation

    ComposioHQ/awesome-claude-skills

    Automate OCR Web Service tasks via Rube MCP (Composio). An agent skill from ComposioHQ/awesome-claude-skills.

    77k GitHub starsUsed in 3 repos~760 tokens
    Productivity & AutomationAuto-check passed
  • Yao OCR Tools

    YaoApp/yao

    Extracts text from images and PDFs, including invoices, receipts, ID cards and tables, with plain text, JSON or Markdown output and a choice of OCR providers.

    8.1k GitHub stars~1.6k tokensUpdated 2 days ago
    Documents & OfficeAuto-check passed
  • Doc OCR

    davepoon/buildwithclaude

    文档文字识别。用户提供 PDF/扫描件/图片(合同、发票、书页、截图),需要提取文字、转成可编辑文本时使用。扫描件自动 OCR(macOS Vision,中英文)。Document OCR: extract editable text from PDFs, scans, and images (contracts, invoices, book pages, screenshots) via…

    3.6k GitHub stars~244 tokensUpdated yesterday
    Documents & OfficeAuto-check passed
  • Image OCR

    aipoch/medical-research-skills

    Extract text from images with Tesseract OCR; use it when you need to recognize text from PNG/JPEG/TIFF/BMP images, select a language model, or run OCR via natural-language requests (e.g., "Interpret…

    2k GitHub stars~1.7k tokensUpdated 20 days ago
    Auto-check passed
  • Post OCR Cleanup

    brycewang-stanford/Auto-Empirical-Research-Skills

    Clean post-OCR text: correction, QA, multilingual handling, provenance.

    4.5k GitHub stars~4.7k tokensUpdated 2 days ago
    Documents & OfficeAuto-check passed
  • Vlm OCR Pipeline

    brycewang-stanford/Auto-Empirical-Research-Skills

    VLM-based OCR pipeline: model selection, prompts, architecture, evaluation.

    4.5k GitHub stars~3.7k tokensUpdated 2 days ago
    Documents & OfficeAuto-check passed

More from countbot-ai/CountBot

All 11 skills in this repo
  • Ima Knowledge Base

    countbot-ai/CountBot

    通过 IMA OpenAPI 处理知识库任务。支持知识库内容搜索、命中详情查看、条目浏览、列出知识库、上传文件、导入网页。用户提到知识库、资料库、上传到知识库、导入网页、搜知识库时使用。

    781 GitHub stars~506 tokensUpdated 3 days ago
    Auto-check passed
  • News

    countbot-ai/CountBot

    新闻与资讯查询。获取中文新闻和全球 AI 技术资讯,支持按分类查询(时政、财经、科技、社会、国际、体育、娱乐、AI 技术、AI 社区)。当用户询问最新新闻、AI 动态、行业资讯时使用。

    781 GitHub starsUsed in 1 repo~876 tokens
    Auto-check passed
  • Web Design

    countbot-ai/CountBot

    网页设计与部署。生成精美的单页 HTML 网页(报告、落地页、数据可视化等),支持一键部署到 Cloudflare Pages。使用 Tailwind CSS + Chart.js + Font Awesome 技术栈。当用户要求制作网页、生成报告页面、创建落地页、数据可视化展示、部署网页到线上时使用。

    781 GitHub stars~1k tokensUpdated 3 days ago
    Auto-check passed
  • Agent Team Manager

    countbot-ai/CountBot

    多智能体团队管理。创建、查看、修改、删除 CountBot 的多智能体团队,管理团队成员(角色)和团队级自定义模型配置。当用户要新建 Pipeline/Graph/Council 团队、调整成员分工、修改依赖关系、开关技能系统、设置团队专属模型时使用。

    781 GitHub stars~1.6k tokensUpdated 3 days ago
    Auto-check passed
  • Cron Manager

    countbot-ai/CountBot

    定时任务管理。创建、查看、修改、删除定时任务,管理任务会话数据。当用户需要设置提醒、定时执行任务、管理调度计划时使用. An agent skill from countbot-ai/CountBot.

    781 GitHub stars~1.4k tokensUpdated 3 days ago
    Auto-check passed
  • Find Skills

    countbot-ai/CountBot

    基于腾讯 SkillHub 搜索、安装和管理技能。用户提到“找技能”“安装 skill”“扩展功能”“启用/禁用 skill”“删除 skill”“安装 SkillHub CLI”时优先使用。

    781 GitHub stars~794 tokensUpdated 3 days ago
    Auto-check passed

Questions about Image Analysis

What does Image Analysis do?

图片分析与识别,可分析本地图片、网络图片、视频、文件。适用于 OCR、物体识别、场景理解等。当用户发送图片或要求分析图片时必须使用此技能。. Image Analysis is an agent skill from countbot-ai/CountBot.

How do I install Image Analysis in Claude Code?

Run `npx skills add countbot-ai/CountBot --skill image-analysis -a claude-code`. Or copy the skill folder (workspace/skills/image-analysis in countbot-ai/CountBot) into .claude/skills/image-analysis in your project. Claude Code loads it when a task matches its description.

How do I install Image Analysis in Codex?

Run `npx skills add countbot-ai/CountBot --skill image-analysis -a codex`. Or copy the skill folder (workspace/skills/image-analysis in countbot-ai/CountBot) into .agents/skills/image-analysis in your project. Codex loads it when a task matches its description.

Can I use Image Analysis in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add countbot-ai/CountBot --skill image-analysis -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/image-analysis, .gemini/skills/image-analysis, .github/skills/image-analysis and .opencode/skills/image-analysis in your project.

What does Image Analysis need to run?

Going by SKILL.md and its folder, Image Analysis needs Python for the scripts in its folder and the command-line tools its instructions call (python3). Our summary lists: Python 3.

Does Image Analysis access the network?

SKILL.md names 2 domains. As links in the text: open.bigmodel.cn and help.aliyun.com. This is read from the text; nothing was executed.

Is Image Analysis safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. The check reads SKILL.md only: the scripts in the folder are not scanned, so read them before running anything.

What licence does Image Analysis use?

Image Analysis is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Image Analysis use?

About 532 tokens (SKILL.md is roughly 2.1k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Image Analysis?

Skills that share tags, products or a category with Image Analysis: OCR Web Service Automation (ComposioHQ/awesome-claude-skills, 77k stars), Yao OCR Tools (YaoApp/yao, 8.1k stars), Doc OCR (davepoon/buildwithclaude, 3.6k stars) and Image OCR (aipoch/medical-research-skills, 2k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Image Analysis?

countbot-ai (a GitHub user) maintains it in countbot-ai/CountBot, which has 781 GitHub stars. The repository holds 11 skills in this directory. The repository was last updated on October 4, 2026.

Source: countbot-ai/CountBot on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.