Iib
zanllp/infinite-image-browsing
Interact with IIB (Infinite Image Browsing) service for searching, browsing, tagging, and organizing AI-generated images.
Deploy and maintain image understanding (OCR + local VLM + cloud VL) and image generation for Codex connected to text-only models like DeepSeek.
$ npx skills add aiskillstore/marketplace --skill deepseek-vision-bridge -a claude-codeProject install by default; add -g for ~/.claude/skills/.
$ gh skill install aiskillstore/marketplace deepseek-vision-bridge --agent claude-codeProject scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).
$ git clone --depth 1 https://github.com/aiskillstore/marketplace.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/lightlw/deepseek-vision-bridge .claude/skills/deepseek-vision-bridge && rm -rf skills-srcUse ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.
Claude Code skills documentation · loads skills from .claude/skills/
Install the "deepseek-vision-bridge" agent skill from https://github.com/aiskillstore/marketplace/tree/main/skills/lightlw/deepseek-vision-bridge into .claude/skills/deepseek-vision-bridge/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "deepseek-vision-bridge", then confirm the skill loads.Claude Code copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$skill-installer install https://github.com/aiskillstore/marketplace/tree/main/skills/lightlw/deepseek-vision-bridgeType this inside Codex. $skill-installer <name> installs a curated skill from openai/skills. The installer writes to $CODEX_HOME/skills (default ~/.codex/skills). Restart Codex if the skill does not show up.
$ npx skills add aiskillstore/marketplace --skill deepseek-vision-bridge -a codexProject install goes to .agents/skills/; add -g for ~/.codex/skills/.
$ gh skill install aiskillstore/marketplace deepseek-vision-bridge --agent codexProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/aiskillstore/marketplace.git skills-src && mkdir -p .agents/skills && cp -r skills-src/skills/lightlw/deepseek-vision-bridge .agents/skills/deepseek-vision-bridge && rm -rf skills-srcUse ~/.agents/skills/ instead of .agents/skills for a personal install.
Codex skills documentation · loads skills from .agents/skills/
Install the "deepseek-vision-bridge" agent skill from https://github.com/aiskillstore/marketplace/tree/main/skills/lightlw/deepseek-vision-bridge into .agents/skills/deepseek-vision-bridge/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "deepseek-vision-bridge", then confirm the skill loads.Codex copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add aiskillstore/marketplace --skill deepseek-vision-bridge -a cursorProject install goes to .agents/skills/; add -g for ~/.cursor/skills/.
$ gh skill install aiskillstore/marketplace deepseek-vision-bridge --agent cursorProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/aiskillstore/marketplace.git skills-src && mkdir -p .cursor/skills && cp -r skills-src/skills/lightlw/deepseek-vision-bridge .cursor/skills/deepseek-vision-bridge && rm -rf skills-srcUse ~/.cursor/skills/ instead of .cursor/skills for a personal install.
Cursor skills documentation · loads skills from .cursor/skills/, .agents/skills/, .claude/skills/, .codex/skills/
Install the "deepseek-vision-bridge" agent skill from https://github.com/aiskillstore/marketplace/tree/main/skills/lightlw/deepseek-vision-bridge into .cursor/skills/deepseek-vision-bridge/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "deepseek-vision-bridge", then confirm the skill loads.Cursor copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gemini skills install https://github.com/aiskillstore/marketplace.git --path skills/lightlw/deepseek-vision-bridge--scope user (default) or --scope workspace; --path is the subfolder of the repo that holds the skill; --consent skips the security confirmation prompt.
$ npx skills add aiskillstore/marketplace --skill deepseek-vision-bridge -a gemini-cliProject install goes to .agents/skills/; add -g for ~/.gemini/skills/.
$ gh skill install aiskillstore/marketplace deepseek-vision-bridge --agent gemini-cliProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/aiskillstore/marketplace.git skills-src && mkdir -p .gemini/skills && cp -r skills-src/skills/lightlw/deepseek-vision-bridge .gemini/skills/deepseek-vision-bridge && rm -rf skills-srcUse ~/.gemini/skills/ instead of .gemini/skills for a personal install, then run /skills reload.
Gemini CLI skills documentation · loads skills from .gemini/skills/, .agents/skills/
Install the "deepseek-vision-bridge" agent skill from https://github.com/aiskillstore/marketplace/tree/main/skills/lightlw/deepseek-vision-bridge into .gemini/skills/deepseek-vision-bridge/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "deepseek-vision-bridge", then confirm the skill loads.Gemini CLI copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gh skill install aiskillstore/marketplace deepseek-vision-bridgeInstalls for Copilot at project scope by default; add --scope user for a personal install. Preview a skill first with gh skill preview. Needs GitHub CLI 2.90.0 or later (public preview).
$ npx skills add aiskillstore/marketplace --skill deepseek-vision-bridge -a github-copilotProject install goes to .agents/skills/; add -g for ~/.copilot/skills/.
$ git clone --depth 1 https://github.com/aiskillstore/marketplace.git skills-src && mkdir -p .github/skills && cp -r skills-src/skills/lightlw/deepseek-vision-bridge .github/skills/deepseek-vision-bridge && rm -rf skills-srcUse ~/.copilot/skills/ instead of .github/skills for a personal install. Commit .github/skills so cloud agent and code review can use it.
GitHub Copilot skills documentation · loads skills from .github/skills/, .claude/skills/, .agents/skills/
Install the "deepseek-vision-bridge" agent skill from https://github.com/aiskillstore/marketplace/tree/main/skills/lightlw/deepseek-vision-bridge into .github/skills/deepseek-vision-bridge/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "deepseek-vision-bridge", then confirm the skill loads.GitHub Copilot copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add aiskillstore/marketplace --skill deepseek-vision-bridge -a opencodeOpenCode documents no install command of its own. Project install goes to .agents/skills/; add -g for ~/.config/opencode/skills/.
$ gh skill install aiskillstore/marketplace deepseek-vision-bridge --agent opencodeProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/aiskillstore/marketplace.git skills-src && mkdir -p .opencode/skills && cp -r skills-src/skills/lightlw/deepseek-vision-bridge .opencode/skills/deepseek-vision-bridge && rm -rf skills-srcUse ~/.config/opencode/skills/ instead of .opencode/skills for a personal install.
OpenCode skills documentation · loads skills from .opencode/skills/, .claude/skills/, .agents/skills/
Install the "deepseek-vision-bridge" agent skill from https://github.com/aiskillstore/marketplace/tree/main/skills/lightlw/deepseek-vision-bridge into .opencode/skills/deepseek-vision-bridge/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "deepseek-vision-bridge", then confirm the skill loads.OpenCode copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
deepseek-vision-bridgeDeploy and maintain image understanding (OCR + local VLM + cloud VL) and image generation for Codex connected to text-only models like DeepSeek.
Deepseek Vision Bridge is an agent skill from aiskillstore/marketplace. Deploy and maintain image understanding (OCR + local VLM + cloud VL) and image generation for Codex connected to text-only models like DeepSeek. Use when the user wants to see/read images in the conversation bar with a text-only model, set up automatic OCR or vision-model processing for pasted images, configure image generation via SiliconFlow or local Stable Diffusion/Flux, fix a broken vision proxy, or asks "看图/生图/贴图识别/图片理解/OCR代理" on a Codex+DeepSeek setup. Deploys a local proxy between Codex and the model that…
Its SKILL.md is about 810 tokens, which your agent loads only when the skill is triggered. The skill folder holds 22 other files, including scripts, reference files and assets (for example `README.md`, `agents/openai.yaml` and `assets/image-gen.js`).
It sits in Media & Creative, covering Image generation, Diffusion and image models and Computer vision. It works with DeepSeek, Stable Diffusion and Ollama. The repository describes itself as: Security-audited skills for Claude, Codex & Claude Code. One-click install, quality verified. The licence is MIT.
7 steps, taken from the step headings in SKILL.md.
Read from SKILL.md and the folder at commit 44923f3. It shows what the files ask for, not the result of running them.
Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.
From allowed-tools in the SKILL.md frontmatter.
Ships 1 file in scripts/ (JavaScript, PowerShell and Shell, from the files we listed), which the agent can run.
Shell commands in SKILL.md call:
pythonollamanpmFrom the folder's file list and the shell code blocks in SKILL.md.
Hosts in commands or code, which the agent is likely to contact:
ollama.comFrom URLs in SKILL.md, links to its own repository left out.
Names these keys or tokens, usually read from environment variables:
DEEPSEEK_API_KEYSILICONFLOW_API_KEYFrom names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.
Deepseek Vision Bridge loads about 814 tokens when it runs, and up to ~2.9k if it reads all its reference files. Until then it costs about 146 tokens; SKILL.md has 149 words of instructions outside code blocks.
Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.
The automated check noted patterns worth knowing about, such as sudo or a known installer.
├── .env ← 从 assets/.env.template 复制并填写复制 assets 下文件到代理目录,然后编辑 `.env`:- 每台机器的 key、端口、模型名不同,以用户机器的 .env / config.toml 为准Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); the scripts in this folder are not scanned.
The full file from aiskillstore/marketplace at commit 44923f3, republished under its MIT licence (© aiskillstore). 149 words, ~814 tokens.
.claude/skills/deepseek-vision-bridge/SKILL.md (or your agent's skills folder). This skill also uses 19 other files; get the full folder from GitHub.为 Codex + 纯文本模型(DeepSeek)部署"看图 + 生图"能力。
核心机制:本地代理拦截 Codex 请求中的图片,用三级引擎(OCR / 本地VLM / 云端VL) 转成文字后再转发给 DeepSeek;生图由 image-gen.js 完成后返回文件路径。
运行 scripts/check_env.py(需要 Python 3),得到 JSON 报告:
python scripts/check_env.py关注字段:node、ollama_installed、ollama_models、gpu、
proxy_port_open、codex_config、deepseek_key_configured。
选定代理目录(例如 ~/codex-vision-bridge/ 或用户偏好位置),创建:
<代理目录>/
├── ocr-proxy.js ← 从 assets/ 复制
├── image-gen.js ← 从 assets/ 复制
├── config.json ← 从 assets/config.json.template 复制
├── package.json ← 从 assets/ 复制
├── .env ← 从 assets/.env.template 复制并填写
├── start-proxy.ps1 / .sh ← 从 assets/ 复制(按平台)
└── node_modules/ ← npm install 生成按 references/model-guide.md 与用户确认路线:
复制 assets 下文件到代理目录,然后编辑 .env:
DEEPSEEK_API_KEY=sk-... # 必填,DeepSeek 平台获取
SILICONFLOW_API_KEY=sk-... # 可选,硅基流动获取
LOCAL_VL_MODEL=minicpm-v:8b # 可选,本地 VLM 模型名
OCR_PROXY_PORT=57323 # 默认若用户选本地 VLM,先安装 Ollama 并拉模型:
# Windows/macOS: https://ollama.com/download
ollama pull minicpm-v:8bcd <代理目录>
npm install # 安装 tesseract.js
./start-proxy.sh # macOS/Linux
# 或 powershell -ExecutionPolicy Bypass -File start-proxy.ps1 # Windows验证端口监听:57323(或自定义端口)。启动日志在 <代理目录>/outputs/proxy.log。
修改 ~/.codex/config.toml(Windows 为 %USERPROFILE%\.codex\config.toml):
model = "deepseek-v4-flash"
model_provider = "custom"
[model_providers.custom]
name = "custom"
wire_api = "responses"
requires_openai_auth = true
base_url = "http://127.0.0.1:57323/v1"
approvals_reviewer = "user"关键:model_provider = "custom" + base_url 指向代理端口。
改配置前先备份原文件。注意 config.toml 用 UTF-8 无 BOM。
powershell -ExecutionPolicy Bypass -File install-autostart.ps1见 references/troubleshooting.md,按症状查:
端口未监听 / 400 / 401 / 超时 / OCR 乱码 / Ollama 连接失败 / 生图失败。
references/model-guide.mdreferences/architecture.md© aiskillstore, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file
SKILL.md and 19 other files (scripts, references, assets) in skills/lightlw/deepseek-vision-bridge of aiskillstore/marketplace.
Open the folder on GitHubat commit 44923f3
Deepseek Vision Bridge next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.
| Skill | Stars | Used in | Tokens | Auto-check | Licence | Repo updated |
|---|---|---|---|---|---|---|
| Deepseek Vision Bridge this skillaiskillstore/marketplace | 433 | — | ~814 | Automated safety check: Notes | MIT | |
| Iibzanllp/infinite-image-browsing | 1.4k | — | ~3.3k | Automated safety check: Pass | MIT | |
| Stable Diffusion with DiffusersOrchestra-Research/AI-Research-SKILLs | 13k | 5 repos | ~3.2k | Automated safety check: Pass | MIT | |
| Anima Baseartokun/comfyui-mcp | 803 | — | ~4k | Automated safety check: Pass | MIT | |
| Imageguaardvark/guaardvark | 257 | — | ~1.8k | Automated safety check: Pass | MIT | |
| AI Image Prompts SkillLeoYeAI/openclaw-master-skills | 2.2k | — | ~4.3k | Automated safety check: Pass | MIT |
zanllp/infinite-image-browsing
Interact with IIB (Infinite Image Browsing) service for searching, browsing, tagging, and organizing AI-generated images.
Orchestra-Research/AI-Research-SKILLs
Generates and edits images with Stable Diffusion through Hugging Face Diffusers, covering text-to-image, image-to-image, inpainting, SDXL and custom pipelines.
artokun/comfyui-mcp
Anime/illustration text-to-image (ANIMA 1.0, ~2B Cosmos DiT).
guaardvark/guaardvark
Generate or edit images on the user's own GPU through Guaardvark: single images, instruction edits, background cut-outs, inpaint and outpaint, consistent characters from the Cast Library, and batch…
LeoYeAI/openclaw-master-skills
Recommend curated prompts from a 10,000+ real-world image generation prompt library.
HuangYuChuh/ComfyUI_Skills_OpenClaw
Run registered ComfyUI workflows through the fast comfyui-skill CLI, and use the official local Comfy MCP for live template, node, model, validation, and orchestration capabilities.
aiskillstore/marketplace
Analyze codebase with tokei (fast line counts by language) and difft (semantic AST-aware diffs).
aiskillstore/marketplace
Process JSON with jq and YAML/TOML with yq. An agent skill from aiskillstore/marketplace.
aiskillstore/marketplace
Scans for project documentation files (AGENTS.md, CLAUDE.md, GEMINI.md, COPILOT.md, CURSOR.md, WARP.md, and 15+ other formats) and synthesizes guidance.
aiskillstore/marketplace
Modern file and content search using fd, ripgrep (rg), and fzf.
aiskillstore/marketplace
Modern find-and-replace using sd (simpler than sed) and batch replacement patterns.
aiskillstore/marketplace
Detects stale project plans and suggests session commands. An agent skill from aiskillstore/marketplace.
Works with
Categories
Deploy and maintain image understanding (OCR + local VLM + cloud VL) and image generation for Codex connected to text-only models like DeepSeek. Deepseek Vision Bridge is an agent skill from aiskillstore/marketplace. Deploy and maintain image understanding (OCR + local VLM + cloud VL) and image generation for Codex connected to text-only models like DeepSeek.
Deepseek Vision Bridge fits situations like: the user wants to see/read images in the conversation bar with a text-only model; set up automatic OCR; vision-model processing for pasted images; configure image generation via SiliconFlow.
Run `npx skills add aiskillstore/marketplace --skill deepseek-vision-bridge -a claude-code`. Or copy the skill folder (skills/lightlw/deepseek-vision-bridge in aiskillstore/marketplace) into .claude/skills/deepseek-vision-bridge in your project. Claude Code loads it when a task matches its description.
Run `npx skills add aiskillstore/marketplace --skill deepseek-vision-bridge -a codex`. Or copy the skill folder (skills/lightlw/deepseek-vision-bridge in aiskillstore/marketplace) into .agents/skills/deepseek-vision-bridge in your project. Codex loads it when a task matches its description.
Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add aiskillstore/marketplace --skill deepseek-vision-bridge -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/deepseek-vision-bridge, .gemini/skills/deepseek-vision-bridge, .github/skills/deepseek-vision-bridge and .opencode/skills/deepseek-vision-bridge in your project.
Going by SKILL.md and its folder, Deepseek Vision Bridge needs JavaScript, PowerShell and a shell for the scripts in its folder, the command-line tools its instructions call (python, ollama and npm) and credentials named DEEPSEEK_API_KEY and SILICONFLOW_API_KEY. Our summary lists: Python 3; Node.js; A Bash shell; PowerShell; A credential in DEEPSEEK_API_KEY.
SKILL.md names 1 domain. In commands or code: ollama.com; the agent is likely to contact it when it follows the instructions. This is read from the text; nothing was executed.
Our automated static check of SKILL.md found notes only (mentions a .env file), nothing it rates as a warning. It is not a guarantee. The check reads SKILL.md only: the scripts in the folder are not scanned, so read them before running anything.
Deepseek Vision Bridge is published under the MIT licence (from the LICENSE file in the skill folder). It allows redistribution, so the full SKILL.md is shown on this page.
About 814 tokens (SKILL.md is roughly 3.3k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full. Its references folder adds about 2.1k tokens, read only when the agent opens those files.
Skills that share tags, products or a category with Deepseek Vision Bridge: Iib (zanllp/infinite-image-browsing, 1.4k stars), Stable Diffusion with Diffusers (Orchestra-Research/AI-Research-SKILLs, 13k stars), Anima Base (artokun/comfyui-mcp, 803 stars) and Image (guaardvark/guaardvark, 257 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.
aiskillstore (a GitHub organization) maintains it in aiskillstore/marketplace, which has 433 GitHub stars. The repository holds 1,044 skills in this directory. The repository was last updated on October 10, 2026.
Source: aiskillstore/marketplace on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.