Markdown Article Formatter
JimLiu/baoyu-skills
Reformats plain text or Markdown articles with frontmatter, a title, a summary, headings, bold, lists and code blocks, and saves a separate formatted copy.
Generate captions (descriptions) for images, videos, and documents using ZhiPu GLM-V multimodal model series.
$ npx skills add zai-org/GLM-skills --skill glmv-caption -a claude-codeProject install by default; add -g for ~/.claude/skills/.
$ gh skill install zai-org/GLM-skills glmv-caption --agent claude-codeProject scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).
$ git clone --depth 1 https://github.com/zai-org/GLM-skills.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/glmv-caption .claude/skills/glmv-caption && rm -rf skills-srcUse ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.
Claude Code skills documentation · loads skills from .claude/skills/
Install the "glmv-caption" agent skill from https://github.com/zai-org/GLM-skills/tree/main/skills/glmv-caption into .claude/skills/glmv-caption/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "glmv-caption", then confirm the skill loads.Claude Code copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$skill-installer install https://github.com/zai-org/GLM-skills/tree/main/skills/glmv-captionType this inside Codex. $skill-installer <name> installs a curated skill from openai/skills. The installer writes to $CODEX_HOME/skills (default ~/.codex/skills). Restart Codex if the skill does not show up.
$ npx skills add zai-org/GLM-skills --skill glmv-caption -a codexProject install goes to .agents/skills/; add -g for ~/.codex/skills/.
$ gh skill install zai-org/GLM-skills glmv-caption --agent codexProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/zai-org/GLM-skills.git skills-src && mkdir -p .agents/skills && cp -r skills-src/skills/glmv-caption .agents/skills/glmv-caption && rm -rf skills-srcUse ~/.agents/skills/ instead of .agents/skills for a personal install.
Codex skills documentation · loads skills from .agents/skills/
Install the "glmv-caption" agent skill from https://github.com/zai-org/GLM-skills/tree/main/skills/glmv-caption into .agents/skills/glmv-caption/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "glmv-caption", then confirm the skill loads.Codex copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add zai-org/GLM-skills --skill glmv-caption -a cursorProject install goes to .agents/skills/; add -g for ~/.cursor/skills/.
$ gh skill install zai-org/GLM-skills glmv-caption --agent cursorProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/zai-org/GLM-skills.git skills-src && mkdir -p .cursor/skills && cp -r skills-src/skills/glmv-caption .cursor/skills/glmv-caption && rm -rf skills-srcUse ~/.cursor/skills/ instead of .cursor/skills for a personal install.
Cursor skills documentation · loads skills from .cursor/skills/, .agents/skills/, .claude/skills/, .codex/skills/
Install the "glmv-caption" agent skill from https://github.com/zai-org/GLM-skills/tree/main/skills/glmv-caption into .cursor/skills/glmv-caption/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "glmv-caption", then confirm the skill loads.Cursor copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gemini skills install https://github.com/zai-org/GLM-skills.git --path skills/glmv-caption--scope user (default) or --scope workspace; --path is the subfolder of the repo that holds the skill; --consent skips the security confirmation prompt.
$ npx skills add zai-org/GLM-skills --skill glmv-caption -a gemini-cliProject install goes to .agents/skills/; add -g for ~/.gemini/skills/.
$ gh skill install zai-org/GLM-skills glmv-caption --agent gemini-cliProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/zai-org/GLM-skills.git skills-src && mkdir -p .gemini/skills && cp -r skills-src/skills/glmv-caption .gemini/skills/glmv-caption && rm -rf skills-srcUse ~/.gemini/skills/ instead of .gemini/skills for a personal install, then run /skills reload.
Gemini CLI skills documentation · loads skills from .gemini/skills/, .agents/skills/
Install the "glmv-caption" agent skill from https://github.com/zai-org/GLM-skills/tree/main/skills/glmv-caption into .gemini/skills/glmv-caption/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "glmv-caption", then confirm the skill loads.Gemini CLI copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gh skill install zai-org/GLM-skills glmv-captionInstalls for Copilot at project scope by default; add --scope user for a personal install. Preview a skill first with gh skill preview. Needs GitHub CLI 2.90.0 or later (public preview).
$ npx skills add zai-org/GLM-skills --skill glmv-caption -a github-copilotProject install goes to .agents/skills/; add -g for ~/.copilot/skills/.
$ git clone --depth 1 https://github.com/zai-org/GLM-skills.git skills-src && mkdir -p .github/skills && cp -r skills-src/skills/glmv-caption .github/skills/glmv-caption && rm -rf skills-srcUse ~/.copilot/skills/ instead of .github/skills for a personal install. Commit .github/skills so cloud agent and code review can use it.
GitHub Copilot skills documentation · loads skills from .github/skills/, .claude/skills/, .agents/skills/
Install the "glmv-caption" agent skill from https://github.com/zai-org/GLM-skills/tree/main/skills/glmv-caption into .github/skills/glmv-caption/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "glmv-caption", then confirm the skill loads.GitHub Copilot copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add zai-org/GLM-skills --skill glmv-caption -a opencodeOpenCode documents no install command of its own. Project install goes to .agents/skills/; add -g for ~/.config/opencode/skills/.
$ gh skill install zai-org/GLM-skills glmv-caption --agent opencodeProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/zai-org/GLM-skills.git skills-src && mkdir -p .opencode/skills && cp -r skills-src/skills/glmv-caption .opencode/skills/glmv-caption && rm -rf skills-srcUse ~/.config/opencode/skills/ instead of .opencode/skills for a personal install.
OpenCode skills documentation · loads skills from .opencode/skills/, .claude/skills/, .agents/skills/
Install the "glmv-caption" agent skill from https://github.com/zai-org/GLM-skills/tree/main/skills/glmv-caption into .opencode/skills/glmv-caption/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "glmv-caption", then confirm the skill loads.OpenCode copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
glmv-captionGenerate captions (descriptions) for images, videos, and documents using ZhiPu GLM-V multimodal model series.
Glmv Caption is an agent skill from zai-org/GLM-skills. Generate captions (descriptions) for images, videos, and documents using ZhiPu GLM-V multimodal model series. Use this skill whenever the user wants to describe, caption, summarize, or interpret the content of images, videos, or files. Supports single/multiple inputs, URLs, local paths, and base64 (images only).
Its SKILL.md is about 2k tokens, which your agent loads only when the skill is triggered. The skill folder holds 2 other files, including scripts (for example `scripts/glmv_caption.py`).
It sits in Documents & Office. It works with Zhipu GLM. The repository describes itself as: Official skills for the GLM family of models. The licence is Apache-2.0.
3 steps, taken from the first numbered list in SKILL.md.
Read from SKILL.md and the folder at commit 2ecd31c. It shows what the files ask for, not the result of running them.
Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.
From allowed-tools in the SKILL.md frontmatter.
Ships 1 file in scripts/ (Python), which the agent can run.
Shell commands in SKILL.md call:
pythonFrom the folder's file list and the shell code blocks in SKILL.md.
Hosts in commands or code, which the agent is likely to contact:
bigmodel.cnAlso links to:
docs.bigmodel.cnFrom URLs in SKILL.md, links to its own repository left out.
Names these keys or tokens, usually read from environment variables:
ZHIPU_API_KEYFrom names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.
Glmv Caption loads about 2k tokens when it runs. Until then it costs about 82 tokens; SKILL.md has 566 words of instructions outside code blocks.
Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.
The automated check noted patterns worth knowing about, such as sudo or a known installer.
3. **.env file / .env 文件:** Create `.env` in this skill directory:Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); the scripts in this folder are not scanned.
The full file from zai-org/GLM-skills at commit 2ecd31c, republished under its Apache-2.0 licence (© zai-org). 566 words, ~2,003 tokens.
.claude/skills/glmv-caption/SKILL.md (or your agent's skills folder). This skill also uses 1 other file; get the full folder from GitHub.Generate captions for images, videos, and documents using the ZhiPu GLM-V multimodal model.
| Type | Formats | Max Size | Max Count | Base64 |
|---|---|---|---|---|
| Image | jpg, png, jpeg | 5MB / 6000×6000px | 50 | ✅ |
| Video | mp4, mkv, mov | 200MB | — | ❌ |
| File | pdf, docx, txt, xlsx, pptx, jsonl | — | 50 | ❌ |
⚠️ file_url cannot mix with image_url or video_url in the same request. ⚠️ Videos and files only support URLs — local paths and base64 are NOT supported (images only).
| Resource | Link |
|---|---|
| Get API Key | https://bigmodel.cn/usercenter/proj-mgmt/apikeys |
| API Docs | Chat Completions / 对话补全 |
This script reads the key from the ZHIPU_API_KEY environment variable and shares it with other Zhipu skills.
脚本通过 ZHIPU_API_KEY 环境变量获取密钥,与其他智谱技能共用同一个 key。
Get Key / 获取 Key: Visit Zhipu Open Platform API Keys / 智谱开放平台 API Keys to create or copy your key.
Setup options / 配置方式(任选一种):
OpenClaw config (recommended) / OpenClaw 配置(推荐): Set in openclaw.json under skills.entries.glmv-caption.env:
"glmv-caption": { "enabled": true, "env": { "ZHIPU_API_KEY": "你的密钥" } }Shell environment variable / Shell 环境变量: Add to ~/.zshrc:
export ZHIPU_API_KEY="你的密钥".env file / .env 文件: Create .env in this skill directory:
ZHIPU_API_KEY=你的密钥⛔ MANDATORY RESTRICTIONS - DO NOT VIOLATE ⛔
python scripts/glmv_caption.pyAfter running the script, you must show the full raw output to the user exactly as returned. Do not summarize, truncate, or only say "generated". Users need the original model output to evaluate quality.
python scripts/glmv_caption.py --images "https://example.com/photo.jpg"
python scripts/glmv_caption.py --images /path/to/photo.pngpython scripts/glmv_caption.py --images img1.jpg img2.png "https://example.com/img3.jpg"python scripts/glmv_caption.py --videos "https://example.com/clip.mp4"python scripts/glmv_caption.py --files "https://example.com/report.pdf"
python scripts/glmv_caption.py --files "https://example.com/doc1.docx" "https://example.com/doc2.txt"python scripts/glmv_caption.py --images photo.jpg --prompt "Describe the architecture style in detail"python scripts/glmv_caption.py --images photo.jpg --output result.jsonpython scripts/glmv_caption.py --images photo.jpg --thinkingpython {baseDir}/scripts/glmv_caption.py (--images IMG [IMG...] | --videos VID [VID...] | --files FILE [FILE...]) [OPTIONS]| Parameter | Required | Description |
|---|---|---|
--images, -i | One of | Image paths or URLs (supports multiple, base64 OK) |
--videos, -v | One of | Video paths or URLs (supports multiple, mp4/mkv/mov) |
--files, -f | One of | Document paths or URLs (supports multiple, pdf/docx/txt/xlsx/pptx/jsonl) |
--prompt, -p | No | Custom prompt (default: "请详细描述这张图片的内容" / "Please describe this image in detail") |
--model, -m | No | Model name (default: glm-4.6v) |
--temperature, -t | No | Sampling temperature 0-1 (default: 0.8) |
--top-p | No | Nucleus sampling 0.01-1.0 (default: 0.6) |
--max-tokens | No | Max output tokens (default: 1024, max 32768) |
--thinking | No | Enable thinking/reasoning mode |
--output, -o | No | Save result JSON to file |
--pretty | No | Pretty-print JSON output |
--stream | No | Enable streaming output |
Note: --images, --videos, and --files are mutually exclusive per API limits.
{
"success": true,
"caption": "A landscape photo showing a mountain range at sunset...",
"usage": {
"prompt_tokens": 128,
"completion_tokens": 256,
"total_tokens": 384
}
}Key fields:
success — whether the request succeededcaption — the generated caption textusage — token usage statisticswarning — present when content was blocked by safety reviewerror — error details on failureAPI key not configured:
ZHIPU_API_KEY not configured. Get your API key at: https://bigmodel.cn/usercenter/proj-mgmt/apikeys→ Show exact error to user, guide them to configure
Authentication failed (401/403): API key invalid/expired → reconfigure
Rate limit (429): Quota exhausted → inform user to wait
File not found: Local file missing → check path
Content filtered: warning field present → content blocked by safety review
© zai-org, Apache-2.0. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file
SKILL.md and 1 other file (scripts) in skills/glmv-caption of zai-org/GLM-skills.
Open the folder on GitHubat commit 2ecd31c
Glmv Caption next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.
| Skill | Stars | Used in | Tokens | Auto-check | Licence | Repo updated |
|---|---|---|---|---|---|---|
| Glmv Caption this skillzai-org/GLM-skills | 475 | — | ~2k | Automated safety check: Notes | Apache-2.0 | |
| Markdown Article FormatterJimLiu/baoyu-skills | 26k | 6 repos | ~3.5k | Automated safety check: Pass | MIT | |
| MarkitdownImCa0/just-laws | 781 | 14 repos | ~3.2k | Automated safety check: Notes | MIT | |
| Obsidian MarkdownAtmosphere/atmosphere | 3.8k | 20 repos | ~1.3k | Automated safety check: Pass | Apache-2.0 | |
| DOCXrvdbreemen/OTGW-firmware | 207 | 33 repos | ~4.3k | Automated safety check: Pass | Proprietary | |
| Gzh Designisjiamu/gzh-design-skill | 3.9k | 1 repos | ~2.2k | Automated safety check: Pass | AGPL-3.0 |
JimLiu/baoyu-skills
Reformats plain text or Markdown articles with frontmatter, a title, a summary, headings, bold, lists and code blocks, and saves a separate formatted copy.
ImCa0/just-laws
Convert files and office documents to Markdown. An agent skill from ImCa0/just-laws.
Atmosphere/atmosphere
Create and edit Obsidian Flavored Markdown with wikilinks, embeds, callouts, properties, and other Obsidian-specific syntax.
rvdbreemen/OTGW-firmware
A skill your agent uses whenever the user wants to create, read, edit, or manipulate Word documents (.docx files).
isjiamu/gzh-design-skill
微信公众号文章排版引擎,将 Markdown 转换为可直接粘贴到公众号编辑器的 HTML。主题风格从 references/theme-index.md 注册的自定义主题库中选取,自动章节编号、关键词下划线标记、引言卡片、目录导航、代码块、图片/GIF、作者签名。支持 Markdown / Word(.docx) / PDF / 纯文本输入(非 Markdown…
HKUDS/DeepTutor
Reads, creates and edits Word .docx files with python-docx, and drops to raw OOXML for tracked changes, comments and byte-exact edits.
zai-org/GLM-skills
Extract text from images using GLM-OCR API. An agent skill from zai-org/GLM-skills.
zai-org/GLM-skills
Official skill for generating high-quality images from text prompts using ZhiPu GLM-Image API.
zai-org/GLM-skills
Official skill for recognizing and extracting mathematical formulas from images and PDFs into LaTeX format using ZhiPu GLM-OCR API.
zai-org/GLM-skills
Official skill for recognizing handwritten text from images using ZhiPu GLM-OCR API.
zai-org/GLM-skills
Official skill for recognizing and extracting tables from images and PDFs into Markdown format using ZhiPu GLM-OCR API.
zai-org/GLM-skills
Write a textual content based on given document(s) and requirements, using ZhiPu GLM-V multimodal model.
Works with
Categories
Generate captions (descriptions) for images, videos, and documents using ZhiPu GLM-V multimodal model series. Glmv Caption is an agent skill from zai-org/GLM-skills. Generate captions (descriptions) for images, videos, and documents using ZhiPu GLM-V multimodal model series.
Glmv Caption fits situations like: the user wants to describe; interpret the content of images.
Run `npx skills add zai-org/GLM-skills --skill glmv-caption -a claude-code`. Or copy the skill folder (skills/glmv-caption in zai-org/GLM-skills) into .claude/skills/glmv-caption in your project. Claude Code loads it when a task matches its description.
Run `npx skills add zai-org/GLM-skills --skill glmv-caption -a codex`. Or copy the skill folder (skills/glmv-caption in zai-org/GLM-skills) into .agents/skills/glmv-caption in your project. Codex loads it when a task matches its description.
Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add zai-org/GLM-skills --skill glmv-caption -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/glmv-caption, .gemini/skills/glmv-caption, .github/skills/glmv-caption and .opencode/skills/glmv-caption in your project.
Going by SKILL.md and its folder, Glmv Caption needs Python for the scripts in its folder, the command-line tools its instructions call (python) and credentials named ZHIPU_API_KEY. Our summary lists: Python 3; A credential in ZHIPU_API_KEY.
SKILL.md names 2 domains. In commands or code: bigmodel.cn; the agent is likely to contact it when it follows the instructions. As links in the text: docs.bigmodel.cn. This is read from the text; nothing was executed.
Our automated static check of SKILL.md found notes only (mentions a .env file), nothing it rates as a warning. It is not a guarantee. The check reads SKILL.md only: the scripts in the folder are not scanned, so read them before running anything.
Glmv Caption is published under the Apache-2.0 licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.
About 2k tokens (SKILL.md is roughly 8k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.
Skills that share tags, products or a category with Glmv Caption: Markdown Article Formatter (JimLiu/baoyu-skills, 26k stars), Markitdown (ImCa0/just-laws, 781 stars), Obsidian Markdown (Atmosphere/atmosphere, 3.8k stars) and DOCX (rvdbreemen/OTGW-firmware, 207 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.
zai-org (a GitHub organization) maintains it in zai-org/GLM-skills, which has 475 GitHub stars. The repository holds 16 skills in this directory. The repository was last updated on April 15, 2026.
Source: zai-org/GLM-skills on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.