Imagen
sanjay3290/ai-skills
Generate images using Google Gemini's image generation capabilities.
基于文档内容自动生成配图。AI 智能分析文档结构,归纳核心要点, 为每个主题生成符合特定风格的配图。支持封面图生成和自定义图片比例。
$ npx skills add op7418/Document-illustrator-skill --skill document-illustrator -a claude-codeProject install by default; add -g for ~/.claude/skills/.
$ gh skill install op7418/Document-illustrator-skill document-illustrator --agent claude-codeProject scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).
Claude Code skills documentation · loads skills from .claude/skills/
Install the "document-illustrator" agent skill from https://github.com/op7418/Document-illustrator-skill/tree/main into .claude/skills/document-illustrator/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "document-illustrator", then confirm the skill loads.Claude Code copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add op7418/Document-illustrator-skill --skill document-illustrator -a codexProject install goes to .agents/skills/; add -g for ~/.codex/skills/.
$ gh skill install op7418/Document-illustrator-skill document-illustrator --agent codexProject scope by default (.agents/skills/); add --scope user for a personal install.
Codex skills documentation · loads skills from .agents/skills/
Install the "document-illustrator" agent skill from https://github.com/op7418/Document-illustrator-skill/tree/main into .agents/skills/document-illustrator/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "document-illustrator", then confirm the skill loads.Codex copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add op7418/Document-illustrator-skill --skill document-illustrator -a cursorProject install goes to .agents/skills/; add -g for ~/.cursor/skills/.
$ gh skill install op7418/Document-illustrator-skill document-illustrator --agent cursorProject scope by default (.agents/skills/); add --scope user for a personal install.
Cursor skills documentation · loads skills from .cursor/skills/, .agents/skills/, .claude/skills/, .codex/skills/
Install the "document-illustrator" agent skill from https://github.com/op7418/Document-illustrator-skill/tree/main into .cursor/skills/document-illustrator/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "document-illustrator", then confirm the skill loads.Cursor copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add op7418/Document-illustrator-skill --skill document-illustrator -a gemini-cliProject install goes to .agents/skills/; add -g for ~/.gemini/skills/.
$ gh skill install op7418/Document-illustrator-skill document-illustrator --agent gemini-cliProject scope by default (.agents/skills/); add --scope user for a personal install.
Gemini CLI skills documentation · loads skills from .gemini/skills/, .agents/skills/
Install the "document-illustrator" agent skill from https://github.com/op7418/Document-illustrator-skill/tree/main into .gemini/skills/document-illustrator/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "document-illustrator", then confirm the skill loads.Gemini CLI copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gh skill install op7418/Document-illustrator-skill document-illustratorInstalls for Copilot at project scope by default; add --scope user for a personal install. Preview a skill first with gh skill preview. Needs GitHub CLI 2.90.0 or later (public preview).
$ npx skills add op7418/Document-illustrator-skill --skill document-illustrator -a github-copilotProject install goes to .agents/skills/; add -g for ~/.copilot/skills/.
GitHub Copilot skills documentation · loads skills from .github/skills/, .claude/skills/, .agents/skills/
Install the "document-illustrator" agent skill from https://github.com/op7418/Document-illustrator-skill/tree/main into .github/skills/document-illustrator/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "document-illustrator", then confirm the skill loads.GitHub Copilot copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add op7418/Document-illustrator-skill --skill document-illustrator -a opencodeOpenCode documents no install command of its own. Project install goes to .agents/skills/; add -g for ~/.config/opencode/skills/.
$ gh skill install op7418/Document-illustrator-skill document-illustrator --agent opencodeProject scope by default (.agents/skills/); add --scope user for a personal install.
OpenCode skills documentation · loads skills from .opencode/skills/, .claude/skills/, .agents/skills/
Install the "document-illustrator" agent skill from https://github.com/op7418/Document-illustrator-skill/tree/main into .opencode/skills/document-illustrator/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "document-illustrator", then confirm the skill loads.OpenCode copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
document-illustrator基于文档内容自动生成配图。AI 智能分析文档结构,归纳核心要点, 为每个主题生成符合特定风格的配图。支持封面图生成和自定义图片比例。
Document Illustrator is an agent skill from op7418/Document-illustrator-skill. 基于文档内容自动生成配图。AI 智能分析文档结构,归纳核心要点, 为每个主题生成符合特定风格的配图。支持封面图生成和自定义图片比例。 使用场景:当用户需要为文档、文章、笔记生成配图时。 关键词:配图、插图、illustration、generate images、document images
Its SKILL.md is about 1.6k tokens, which your agent loads only when the skill is triggered. The skill folder holds 12 other files, including scripts (for example `README.md`, `examples/README.md` and `scripts/generate_illustrations.py`).
It sits in Media & Creative, covering Image generation and Icons and illustration. It works with Adobe Illustrator. The repository describes itself as: 帮你从文档生成对应的多张配图,内置了歸藏精心探索的图片风格,支持 16:9 和 3:4 两种比例,方便发小红书以及推特。 The licence is MIT.
11 steps, taken from the step headings in SKILL.md.
Read from SKILL.md and the folder at commit 8344815. It shows what the files ask for, not the result of running them.
Pre-approves these tools, so the agent can use them without asking each time:
ReadWriteBash(python:*)GlobAskUserQuestionFrom allowed-tools in the SKILL.md frontmatter.
Ships 2 files in scripts/ (Python), which the agent can run.
Shell commands in SKILL.md call:
pipFrom the folder's file list and the shell code blocks in SKILL.md.
Links to these hosts (documentation or services it may open):
makersuite.google.comFrom URLs in SKILL.md, links to its own repository left out.
Names these keys or tokens, usually read from environment variables:
GEMINI_API_KEYFrom names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.
Document Illustrator loads about 1.6k tokens when it runs. Until then it costs about 42 tokens; SKILL.md has 378 words of instructions outside code blocks.
Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.
The automated check noted patterns worth knowing about, such as sudo or a known installer.
在 `~/.claude/skills/document-illustrator/.env` 中配置1. 检查 `.env` 文件中的 `GEMINI_API_KEY`Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); the scripts in this folder are not scanned.
The full file from op7418/Document-illustrator-skill at commit 8344815, republished under its MIT licence (© op7418). 378 words, ~1,588 tokens.
.claude/skills/document-illustrator/SKILL.md (or your agent's skills folder). This skill also uses 9 other files; get the full folder from GitHub.基于 AI 智能分析的文档配图生成工具。无需依赖特定格式,自动理解内容并生成专业配图。
帮我为这个文档生成配图:/path/to/document.md或者:
我想为这篇文章生成一些配图当你请求生成配图时,Claude 会:
无需担心文档格式:
Claude 会询问你的偏好:
请选择图片比例:
1. 16:9 (横屏) - 适合演示文稿、幻灯片、横屏展示
2. 3:4 (竖屏) - 适合社交媒体、手机查看、海报
请选择 (1/2):是否生成封面图?
封面图将概括文档的所有核心信息,作为系列配图的引导。
1. 是 - 生成封面图 + 内容配图
2. 否 - 仅生成内容配图
请选择 (1/2):期望生成多少张内容配图?
建议范围:3-10 张
根据文档内容,推荐生成 6 张
请输入数字:根据你指定的数量,Claude 会智能归纳文档,然后展示给你确认:
📋 内容归纳完成
📄 封面图内容:(如果选择生成)
"AI 编程工具概念演化:从 Rules 到 Skills"
- 核心概念:静态上下文 vs 动态上下文
- 演化路径:Rules → Commands → MCP → Modes → Skills
- 最佳实践:简化为两个核心工具
📚 内容配图(共 6 张):
1. Rules 的诞生与演化
包含:早期模型幻觉问题、rules 文件的作用、静态上下文概念
2. Commands 和工作流打包
包含:固定工作流的出现、slash command、团队分享
3. MCP Servers 带来动态能力
包含:第三方工具集成、OAuth 认证、上下文膨胀问题
4. Modes 和 Subagents 的登场
包含:人设提示词、系统提示词修改、可靠性设计、Hooks 确定性
5. Skills 统一动态上下文
包含:Skills 概念、动态加载、编程工具优化
6. 最佳实践与未来展望
包含:Rules 使用建议、Skills 探索、核心理念总结
✓ 所有内容已覆盖,无遗漏
确认开始生成配图吗?(Y/N)关键保证:
确认后,Claude 调用 Python 脚本生成图片:
🖼️ 开始生成配图...
正在生成封面图...
✓ 已保存: /path/to/document/images/cover.png
正在生成第 1/6 张...
标题: Rules 的诞生与演化
✓ 已保存: /path/to/document/images/illustration-01.png
正在生成第 2/6 张...
标题: Commands 和工作流打包
✓ 已保存: /path/to/document/images/illustration-02.png
...
✨ 完成!共生成 7 张配图(1 张封面 + 6 张内容)输出位置:文档所在目录下的 images/ 文件夹
/path/to/your/document/
└── images/
├── cover.png # 封面图(如果选择生成)
├── illustration-01.png # 第 1 张内容配图
├── illustration-02.png # 第 2 张内容配图
├── illustration-03.png
├── illustration-04.png
├── illustration-05.png
└── illustration-06.png特点:
适用场景:
特点:
适用场景:
特点:
适用场景:
用户: 帮我为 ~/blog/ai-tools.md 生成配图,用票据风格
Claude: 好的,让我先读取这个文档...
[读取文档,分析内容]
检测到文档长度约 3000 字,内容涉及 AI 工具演化。
请选择图片比例:
1. 16:9 (横屏)
2. 3:4 (竖屏)
用户: 1
Claude: 是否生成封面图?
1. 是
2. 否
用户: 是
Claude: 期望生成多少张内容配图?
建议范围:3-10 张
根据文档内容,推荐生成 5 张
用户: 5
Claude: [展示归纳的 5 个主题]
确认开始生成配图吗?
用户: Y
Claude: [调用 Python 脚本生成 6 张图片]
✨ 完成!配图已保存到 ~/blog/images/用户: 为这个纯文本笔记生成配图
Claude: [读取笔记,即使没有任何标题也能理解内容]
[智能识别 3 个核心主题]
[询问配置、展示归纳、生成配图]用户: 我想要 8 张图,竖屏的,不要封面
Claude: 好的,16:9 竖屏,不生成封面,8 张内容配图。
[将文档归纳为 8 个主题]
[展示归纳结果,用户确认后生成]| 比例 | 分辨率 | 适用场景 |
|---|---|---|
| 16:9 | 2560x1440 (2K) | 演示文稿、横屏展示 |
| 3:4 | 1920x2560 | 社交媒体、竖屏查看 |
可选 4K 分辨率:
必需:
pip install google-genai pillow python-dotenvAPI 密钥:
~/.claude/skills/document-illustrator/.env 中配置GEMINI_API_KEYClaude 归纳内容时遵循以下原则:
错误信息:
Error: Invalid API key解决方案:
.env 文件中的 GEMINI_API_KEY问题:归纳的主题不符合预期
解决方案:
可能原因:
解决方案:
| 图片数量 | API 调用次数 | 预估成本 |
|---|---|---|
| 无封面 + 3 张 | 3 次 | 低 |
| 有封面 + 5 张 | 6 次 | 中 |
| 有封面 + 10 张 | 11 次 | 较高 |
建议:
太少:
太多:
推荐:
16:9 适合:
3:4 适合:
建议生成封面图的场景:
可以不生成的场景:
技术文档 → 渐变玻璃卡片风格 数据报告 → 票据风格 教程故事 → 矢量插画风格 产品介绍 → 渐变玻璃卡片风格
[代码] 读取文档 → 识别 ## ### 标题 → 机械切分
↓
依赖特定格式
容易遗漏内容
不够智能[Claude] 读取文档 → AI 理解内容 → 智能归纳主题
↓
格式无关
内容完整
用户可控核心区别:
| 功能 | Document Illustrator | 传统 PPT 工具 | AI 图片生成器 |
|---|---|---|---|
| 理解文档内容 | ✅ AI 智能理解 | ❌ 需要手动 | ❌ 需要手动输入 |
| 格式依赖 | ✅ 格式无关 | ❌ 依赖特定格式 | ✅ 无依赖 |
| 内容完整性 | ✅ 自动验证 | ⚠️ 手动确保 | ❌ 无法保证 |
| 批量生成 | ✅ 一次生成多张 | ❌ 逐张制作 | ⚠️ 需要多次输入 |
| 风格一致性 | ✅ 自动保持 | ⚠️ 手动调整 | ⚠️ 需要重复提示词 |
如有问题或建议:
~/.claude/plans/shimmering-tickling-seahorse.md~/.claude/skills/document-illustrator/让 AI 帮你理解和归纳内容,生成专业配图! ✨
© op7418, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file
SKILL.md and 9 other files (scripts) in the repository root of op7418/Document-illustrator-skill.
Open the folder on GitHubat commit 8344815
Document Illustrator next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.
| Skill | Stars | Used in | Tokens | Auto-check | Licence | Repo updated |
|---|---|---|---|---|---|---|
| Document Illustrator this skillop7418/Document-illustrator-skill | 599 | — | ~1.6k | Automated safety check: Notes | MIT | |
| Imagensanjay3290/ai-skills | 430 | 7 repos | ~657 | Automated safety check: Pass | Apache-2.0 | |
| Canghe Article Illustratorfreestylefly/canghe-skills | 461 | 2 repos | ~1.3k | Automated safety check: Pass | None | |
| Cursor Image Generationtmcfarlane/oh-my-cursor | 109 | — | ~1.8k | Automated safety check: Pass | MIT | |
| Imagegennexu-io/open-design | 100k | — | ~300 | Automated safety check: Pass | Apache-2.0 | |
| Imagennexu-io/open-design | 100k | — | ~279 | Automated safety check: Pass | Apache-2.0 |
sanjay3290/ai-skills
Generate images using Google Gemini's image generation capabilities.
freestylefly/canghe-skills
Analyzes article structure, identifies positions requiring visual aids, generates illustrations with Type × Style two-dimension approach.
tmcfarlane/oh-my-cursor
Generate and iterate images in Cursor using the built-in image model and strong prompts.
nexu-io/open-design
Generate and edit images using OpenAI's Image API for project assets — UI mockups, icons, illustrations, social cards, and visual references.
nexu-io/open-design
Generate images using Google Gemini's image generation API for UI mockups, icons, illustrations, and visual assets.
xiaohuailabs/xiaohu-ip-studio
Creates body illustrations for Chinese long-form articles starring a fixed IP character, with a method of anchor, metaphor and self-check and a library of 31 characters.
Works with
Categories
基于文档内容自动生成配图。AI 智能分析文档结构,归纳核心要点, 为每个主题生成符合特定风格的配图。支持封面图生成和自定义图片比例。. Document Illustrator is an agent skill from op7418/Document-illustrator-skill.
Document Illustrator fits situations like: tasks that involve Image generation; tasks that involve Icons and illustration.
Run `npx skills add op7418/Document-illustrator-skill --skill document-illustrator -a claude-code`. Or copy the skill folder (the op7418/Document-illustrator-skill repository) into .claude/skills/document-illustrator in your project. Claude Code loads it when a task matches its description.
Run `npx skills add op7418/Document-illustrator-skill --skill document-illustrator -a codex`. Or copy the skill folder (the op7418/Document-illustrator-skill repository) into .agents/skills/document-illustrator in your project. Codex loads it when a task matches its description.
Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add op7418/Document-illustrator-skill --skill document-illustrator -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/document-illustrator, .gemini/skills/document-illustrator, .github/skills/document-illustrator and .opencode/skills/document-illustrator in your project.
Going by SKILL.md and its folder, Document Illustrator needs Python for the scripts in its folder, the command-line tools its instructions call (pip) and credentials named GEMINI_API_KEY. Our summary lists: Python 3; A credential in GEMINI_API_KEY. Its frontmatter pre-approves these tools: Read, Write, Bash(python:*), Glob, AskUserQuestion.
SKILL.md names 1 domain. As links in the text: makersuite.google.com. This is read from the text; nothing was executed.
Our automated static check of SKILL.md found notes only (mentions a .env file), nothing it rates as a warning. It is not a guarantee. The check reads SKILL.md only: the scripts in the folder are not scanned, so read them before running anything.
Document Illustrator is published under the MIT licence (from the LICENSE file in the skill folder). It allows redistribution, so the full SKILL.md is shown on this page.
About 1.6k tokens (SKILL.md is roughly 6.4k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.
Skills that share tags, products or a category with Document Illustrator: Imagen (sanjay3290/ai-skills, 430 stars), Canghe Article Illustrator (freestylefly/canghe-skills, 461 stars), Cursor Image Generation (tmcfarlane/oh-my-cursor, 109 stars) and Imagegen (nexu-io/open-design, 100k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.
op7418 (a GitHub user) maintains it in op7418/Document-illustrator-skill, which has 599 GitHub stars. The repository was last updated on January 21, 2026.
Source: op7418/Document-illustrator-skill on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.