OCR Web Service Automation
ComposioHQ/awesome-claude-skills
Automate OCR Web Service tasks via Rube MCP (Composio). An agent skill from ComposioHQ/awesome-claude-skills.
图片分析与识别,可分析本地图片、网络图片、视频、文件。适用于 OCR、物体识别、场景理解等。当用户发送图片或要求分析图片时必须使用此技能。
$ npx skills add countbot-ai/CountBot --skill image-analysis -a claude-codeProject install by default; add -g for ~/.claude/skills/.
$ gh skill install countbot-ai/CountBot image-analysis --agent claude-codeProject scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).
$ git clone --depth 1 https://github.com/countbot-ai/CountBot.git skills-src && mkdir -p .claude/skills && cp -r skills-src/workspace/skills/image-analysis .claude/skills/image-analysis && rm -rf skills-srcUse ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.
Claude Code skills documentation · loads skills from .claude/skills/
Install the "image-analysis" agent skill from https://github.com/countbot-ai/CountBot/tree/main/workspace/skills/image-analysis into .claude/skills/image-analysis/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "image-analysis", then confirm the skill loads.Claude Code copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$skill-installer install https://github.com/countbot-ai/CountBot/tree/main/workspace/skills/image-analysisType this inside Codex. $skill-installer <name> installs a curated skill from openai/skills. The installer writes to $CODEX_HOME/skills (default ~/.codex/skills). Restart Codex if the skill does not show up.
$ npx skills add countbot-ai/CountBot --skill image-analysis -a codexProject install goes to .agents/skills/; add -g for ~/.codex/skills/.
$ gh skill install countbot-ai/CountBot image-analysis --agent codexProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/countbot-ai/CountBot.git skills-src && mkdir -p .agents/skills && cp -r skills-src/workspace/skills/image-analysis .agents/skills/image-analysis && rm -rf skills-srcUse ~/.agents/skills/ instead of .agents/skills for a personal install.
Codex skills documentation · loads skills from .agents/skills/
Install the "image-analysis" agent skill from https://github.com/countbot-ai/CountBot/tree/main/workspace/skills/image-analysis into .agents/skills/image-analysis/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "image-analysis", then confirm the skill loads.Codex copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add countbot-ai/CountBot --skill image-analysis -a cursorProject install goes to .agents/skills/; add -g for ~/.cursor/skills/.
$ gh skill install countbot-ai/CountBot image-analysis --agent cursorProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/countbot-ai/CountBot.git skills-src && mkdir -p .cursor/skills && cp -r skills-src/workspace/skills/image-analysis .cursor/skills/image-analysis && rm -rf skills-srcUse ~/.cursor/skills/ instead of .cursor/skills for a personal install.
Cursor skills documentation · loads skills from .cursor/skills/, .agents/skills/, .claude/skills/, .codex/skills/
Install the "image-analysis" agent skill from https://github.com/countbot-ai/CountBot/tree/main/workspace/skills/image-analysis into .cursor/skills/image-analysis/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "image-analysis", then confirm the skill loads.Cursor copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gemini skills install https://github.com/countbot-ai/CountBot.git --path workspace/skills/image-analysis--scope user (default) or --scope workspace; --path is the subfolder of the repo that holds the skill; --consent skips the security confirmation prompt.
$ npx skills add countbot-ai/CountBot --skill image-analysis -a gemini-cliProject install goes to .agents/skills/; add -g for ~/.gemini/skills/.
$ gh skill install countbot-ai/CountBot image-analysis --agent gemini-cliProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/countbot-ai/CountBot.git skills-src && mkdir -p .gemini/skills && cp -r skills-src/workspace/skills/image-analysis .gemini/skills/image-analysis && rm -rf skills-srcUse ~/.gemini/skills/ instead of .gemini/skills for a personal install, then run /skills reload.
Gemini CLI skills documentation · loads skills from .gemini/skills/, .agents/skills/
Install the "image-analysis" agent skill from https://github.com/countbot-ai/CountBot/tree/main/workspace/skills/image-analysis into .gemini/skills/image-analysis/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "image-analysis", then confirm the skill loads.Gemini CLI copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gh skill install countbot-ai/CountBot image-analysisInstalls for Copilot at project scope by default; add --scope user for a personal install. Preview a skill first with gh skill preview. Needs GitHub CLI 2.90.0 or later (public preview).
$ npx skills add countbot-ai/CountBot --skill image-analysis -a github-copilotProject install goes to .agents/skills/; add -g for ~/.copilot/skills/.
$ git clone --depth 1 https://github.com/countbot-ai/CountBot.git skills-src && mkdir -p .github/skills && cp -r skills-src/workspace/skills/image-analysis .github/skills/image-analysis && rm -rf skills-srcUse ~/.copilot/skills/ instead of .github/skills for a personal install. Commit .github/skills so cloud agent and code review can use it.
GitHub Copilot skills documentation · loads skills from .github/skills/, .claude/skills/, .agents/skills/
Install the "image-analysis" agent skill from https://github.com/countbot-ai/CountBot/tree/main/workspace/skills/image-analysis into .github/skills/image-analysis/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "image-analysis", then confirm the skill loads.GitHub Copilot copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add countbot-ai/CountBot --skill image-analysis -a opencodeOpenCode documents no install command of its own. Project install goes to .agents/skills/; add -g for ~/.config/opencode/skills/.
$ gh skill install countbot-ai/CountBot image-analysis --agent opencodeProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/countbot-ai/CountBot.git skills-src && mkdir -p .opencode/skills && cp -r skills-src/workspace/skills/image-analysis .opencode/skills/image-analysis && rm -rf skills-srcUse ~/.config/opencode/skills/ instead of .opencode/skills for a personal install.
OpenCode skills documentation · loads skills from .opencode/skills/, .claude/skills/, .agents/skills/
Install the "image-analysis" agent skill from https://github.com/countbot-ai/CountBot/tree/main/workspace/skills/image-analysis into .opencode/skills/image-analysis/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "image-analysis", then confirm the skill loads.OpenCode copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
image-analysis图片分析与识别,可分析本地图片、网络图片、视频、文件。适用于 OCR、物体识别、场景理解等。当用户发送图片或要求分析图片时必须使用此技能。
Image Analysis is an agent skill from countbot-ai/CountBot. 图片分析与识别,可分析本地图片、网络图片、视频、文件。适用于 OCR、物体识别、场景理解等。当用户发送图片或要求分析图片时必须使用此技能。
Its SKILL.md is about 530 tokens, which your agent loads only when the skill is triggered. The skill folder holds 5 other files, including scripts (for example `scripts/config.json`, `scripts/vision.py` and `scripts/vision_manager.py`).
The repository describes itself as: 更适配中文用户的轻量开源AI Agent | 国产大模型Coding plan支持 | 兼容OpenClaw Skills生态| 已接入微信ClawBot/微博龙虾/飞书/钉钉/QQ/小智AI/Telegram/deepseek-v4。 The licence is MIT.
Read from SKILL.md and the folder at commit 3c26f11. It shows what the files ask for, not the result of running them.
Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.
From allowed-tools in the SKILL.md frontmatter.
Ships 4 files in scripts/ (Python), which the agent can run.
Shell commands in SKILL.md call:
python3From the folder's file list and the shell code blocks in SKILL.md.
Links to these hosts (documentation or services it may open):
open.bigmodel.cnhelp.aliyun.comFrom URLs in SKILL.md, links to its own repository left out.
Names no API keys, tokens, secrets or passwords.
From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.
Image Analysis loads about 532 tokens when it runs. Until then it costs about 21 tokens; SKILL.md has 43 words of instructions outside code blocks.
Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.
The automated check found no risky patterns in SKILL.md.
Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); the scripts in this folder are not scanned.
The full file from countbot-ai/CountBot at commit 3c26f11, republished under its MIT licence (© countbot-ai). 43 words, ~532 tokens.
.claude/skills/image-analysis/SKILL.md (or your agent's skills folder). This skill also uses 4 other files; get the full folder from GitHub.支持智谱 GLM-4V 和千问 Qwen-VL 两种视觉模型。
当用户发送图片或要求分析图片时,必须使用此技能,不要使用 PIL、pytesseract 等其他方法。
编辑 skills/image-analysis/scripts/config.json:
{
"default_model": "zhipu",
"zhipu": {
"api_key": "your-zhipu-api-key",
"model": "glm-4.6v-flash"
},
"qwen": {
"api_key": "your-qwen-api-key",
"model": "qwen3-vl-plus"
}
}API Key 获取:
# 分析本地图片(最常用)
python3 skills/image-analysis/scripts/vision.py analyze --image 图片路径 --prompt "描述图片内容"
# 分析网络图片
python3 skills/image-analysis/scripts/vision.py analyze --image https://example.com/image.jpg --prompt "描述图片"
# 多图对比
python3 skills/image-analysis/scripts/vision.py analyze --image img1.jpg --image img2.jpg --prompt "对比差异"
# 指定模型
python3 skills/image-analysis/scripts/vision.py analyze --image image.jpg --prompt "描述图片" --model qwen
# 开启思考模式(仅智谱,提升准确度)
python3 skills/image-analysis/scripts/vision.py analyze --image image.jpg --prompt "详细分析" --thinking
# 视频分析
python3 skills/image-analysis/scripts/vision.py analyze --video video.mp4 --prompt "总结视频内容"
# JSON 输出
python3 skills/image-analysis/scripts/vision.py analyze --image image.jpg --prompt "描述图片" --json用户发送图片后,系统下载到本地(如 data/temp/images/xxx.jpg):
# 图片描述
python3 skills/image-analysis/scripts/vision.py analyze --image data/temp/images/xxx.jpg --prompt "描述这张图片的内容"
# OCR 识别
python3 skills/image-analysis/scripts/vision.py analyze --image data/temp/images/xxx.jpg --prompt "提取图片中的所有文字信息"
# 物体定位(开启思考模式)
python3 skills/image-analysis/scripts/vision.py analyze --image data/temp/images/xxx.jpg --prompt "找出物体位置,返回坐标" --thinking| 场景 | 推荐 |
|---|---|
| 简单描述 | 任意 |
| 复杂推理、物体定位 | 智谱 + --thinking |
| 高精度识别、文档解析 | 千问 |
| 成本敏感 | 智谱(免费) |
© countbot-ai, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file
SKILL.md and 4 other files (scripts) in workspace/skills/image-analysis of countbot-ai/CountBot.
Open the folder on GitHubat commit 3c26f11
We found 2 copies of this SKILL.md (exact, near-identical or edited) in other folders, from 1 other GitHub owner. This page covers the copy in countbot-ai/CountBot, which our catalogue first saw on October 7, 2026.
Image Analysis next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.
| Skill | Stars | Used in | Tokens | Auto-check | Licence | Repo updated |
|---|---|---|---|---|---|---|
| Image Analysis this skillcountbot-ai/CountBot | 781 | 1 repos | ~532 | Automated safety check: Pass | MIT | |
| OCR Web Service AutomationComposioHQ/awesome-claude-skills | 77k | 3 repos | ~760 | Automated safety check: Pass | None | |
| Yao OCR ToolsYaoApp/yao | 8.1k | — | ~1.6k | Automated safety check: Pass | Custom licence | |
| Doc OCRdavepoon/buildwithclaude | 3.6k | — | ~244 | Automated safety check: Pass | MIT | |
| Image OCRaipoch/medical-research-skills | 2k | — | ~1.7k | Automated safety check: Pass | MIT | |
| Post OCR Cleanupbrycewang-stanford/Auto-Empirical-Research-Skills | 4.5k | — | ~4.7k | Automated safety check: Pass | Custom licence |
ComposioHQ/awesome-claude-skills
Automate OCR Web Service tasks via Rube MCP (Composio). An agent skill from ComposioHQ/awesome-claude-skills.
YaoApp/yao
Extracts text from images and PDFs, including invoices, receipts, ID cards and tables, with plain text, JSON or Markdown output and a choice of OCR providers.
davepoon/buildwithclaude
文档文字识别。用户提供 PDF/扫描件/图片(合同、发票、书页、截图),需要提取文字、转成可编辑文本时使用。扫描件自动 OCR(macOS Vision,中英文)。Document OCR: extract editable text from PDFs, scans, and images (contracts, invoices, book pages, screenshots) via…
aipoch/medical-research-skills
Extract text from images with Tesseract OCR; use it when you need to recognize text from PNG/JPEG/TIFF/BMP images, select a language model, or run OCR via natural-language requests (e.g., "Interpret…
brycewang-stanford/Auto-Empirical-Research-Skills
Clean post-OCR text: correction, QA, multilingual handling, provenance.
brycewang-stanford/Auto-Empirical-Research-Skills
VLM-based OCR pipeline: model selection, prompts, architecture, evaluation.
countbot-ai/CountBot
通过 IMA OpenAPI 处理知识库任务。支持知识库内容搜索、命中详情查看、条目浏览、列出知识库、上传文件、导入网页。用户提到知识库、资料库、上传到知识库、导入网页、搜知识库时使用。
countbot-ai/CountBot
新闻与资讯查询。获取中文新闻和全球 AI 技术资讯,支持按分类查询(时政、财经、科技、社会、国际、体育、娱乐、AI 技术、AI 社区)。当用户询问最新新闻、AI 动态、行业资讯时使用。
countbot-ai/CountBot
网页设计与部署。生成精美的单页 HTML 网页(报告、落地页、数据可视化等),支持一键部署到 Cloudflare Pages。使用 Tailwind CSS + Chart.js + Font Awesome 技术栈。当用户要求制作网页、生成报告页面、创建落地页、数据可视化展示、部署网页到线上时使用。
countbot-ai/CountBot
多智能体团队管理。创建、查看、修改、删除 CountBot 的多智能体团队,管理团队成员(角色)和团队级自定义模型配置。当用户要新建 Pipeline/Graph/Council 团队、调整成员分工、修改依赖关系、开关技能系统、设置团队专属模型时使用。
countbot-ai/CountBot
定时任务管理。创建、查看、修改、删除定时任务,管理任务会话数据。当用户需要设置提醒、定时执行任务、管理调度计划时使用. An agent skill from countbot-ai/CountBot.
countbot-ai/CountBot
基于腾讯 SkillHub 搜索、安装和管理技能。用户提到“找技能”“安装 skill”“扩展功能”“启用/禁用 skill”“删除 skill”“安装 SkillHub CLI”时优先使用。
图片分析与识别,可分析本地图片、网络图片、视频、文件。适用于 OCR、物体识别、场景理解等。当用户发送图片或要求分析图片时必须使用此技能。. Image Analysis is an agent skill from countbot-ai/CountBot.
Run `npx skills add countbot-ai/CountBot --skill image-analysis -a claude-code`. Or copy the skill folder (workspace/skills/image-analysis in countbot-ai/CountBot) into .claude/skills/image-analysis in your project. Claude Code loads it when a task matches its description.
Run `npx skills add countbot-ai/CountBot --skill image-analysis -a codex`. Or copy the skill folder (workspace/skills/image-analysis in countbot-ai/CountBot) into .agents/skills/image-analysis in your project. Codex loads it when a task matches its description.
Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add countbot-ai/CountBot --skill image-analysis -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/image-analysis, .gemini/skills/image-analysis, .github/skills/image-analysis and .opencode/skills/image-analysis in your project.
Going by SKILL.md and its folder, Image Analysis needs Python for the scripts in its folder and the command-line tools its instructions call (python3). Our summary lists: Python 3.
SKILL.md names 2 domains. As links in the text: open.bigmodel.cn and help.aliyun.com. This is read from the text; nothing was executed.
Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. The check reads SKILL.md only: the scripts in the folder are not scanned, so read them before running anything.
Image Analysis is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.
About 532 tokens (SKILL.md is roughly 2.1k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.
Skills that share tags, products or a category with Image Analysis: OCR Web Service Automation (ComposioHQ/awesome-claude-skills, 77k stars), Yao OCR Tools (YaoApp/yao, 8.1k stars), Doc OCR (davepoon/buildwithclaude, 3.6k stars) and Image OCR (aipoch/medical-research-skills, 2k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.
countbot-ai (a GitHub user) maintains it in countbot-ai/CountBot, which has 781 GitHub stars. The repository holds 11 skills in this directory. The repository was last updated on October 4, 2026.
Source: countbot-ai/CountBot on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.