Agent skill

Glmv Prompt Gen

by zai-org in zai-org/GLM-skills

Analyze images/videos and generate professional prompts for text-to-image and text-to-video AI tools (Midjourney, Stable Diffusion, DALL-E, Sora, Runway, Kling, Pika).

Apache-2.0Auto-check passedMedia & Creative

Install Glmv Prompt Gen

skills CLI
$ npx skills add zai-org/GLM-skills --skill glmv-prompt-gen -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install zai-org/GLM-skills glmv-prompt-gen --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/zai-org/GLM-skills.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/glmv-prompt-gen .claude/skills/glmv-prompt-gen && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
glmv-prompt-gen
GitHub stars
476
Token cost
~1.8k tokens
SKILL.md length
489 words
Files
2 (incl. scripts)
Skills in repo
16
Repo updated
First seen
Licence
Apache-2.0

At a glance

Analyze images/videos and generate professional prompts for text-to-image and text-to-video AI tools (Midjourney, Stable Diffusion, DALL-E, Sora, Runway, Kling, Pika).

  • Works in 2 steps: OpenClaw config (recommended) / OpenClaw… → Shell environment variable / Shell 环境变量:…
  • The user wants to generate prompts from reference images/videos
  • SKILL.md covers When to Use, Supported Input Types, Output Modes and Resource Links, plus 5 more sections
  • Runs Python scripts from its folder; calls python; needs ZHIPU_API_KEY

What it does

Glmv Prompt Gen is an agent skill from zai-org/GLM-skills. Analyze images/videos and generate professional prompts for text-to-image and text-to-video AI tools (Midjourney, Stable Diffusion, DALL-E, Sora, Runway, Kling, Pika). Use when the user wants to generate prompts from reference images/videos, create AI art prompts, or get prompt engineering suggestions from visual content.

Its SKILL.md is about 1.8k tokens, which your agent loads only when the skill is triggered. The skill folder holds 2 other files, including scripts (for example `scripts/prompt_gen.py`).

It sits in Media & Creative, covering AI video generation and Image generation. It works with Midjourney, OpenAI and Stable Diffusion. The repository describes itself as: Official skills for the GLM family of models. The licence is Apache-2.0.

When your agent uses it

  • The user wants to generate prompts from reference images/videos
  • Create AI art prompts
  • Get prompt engineering suggestions from visual content

Example prompts

  • “/glmv-prompt-gen”

Requirements

  • Python 3
  • A credential in ZHIPU_API_KEY

Workflow steps

2 steps, taken from the first numbered list in SKILL.md.

  1. OpenClaw config (recommended) / OpenClaw 配置(推荐): Set in openclaw.json under skills.entries.glmv-prompt-gen.env
  2. Shell environment variable / Shell 环境变量: Add to ~/.zshrc

What it can do on your machine

Read from SKILL.md and the folder at commit 2ecd31c. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Ships 1 file in scripts/ (Python), which the agent can run.

    Shell commands in SKILL.md call:

    • python

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    Links to these hosts (documentation or services it may open):

    • bigmodel.cn
    • docs.bigmodel.cn

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names these keys or tokens, usually read from environment variables:

    • ZHIPU_API_KEY

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Glmv Prompt Gen loads about 1.8k tokens when it runs. Until then it costs about 85 tokens; SKILL.md has 489 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~85
When it runs · the whole SKILL.md, loaded when a task matches
~1.8k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); the scripts in this folder are not scanned.

SKILL.md

The full file from zai-org/GLM-skills at commit 2ecd31c, republished under its Apache-2.0 licence (© zai-org). 489 words, ~1,844 tokens.

Download SKILL.mdSave it as .claude/skills/glmv-prompt-gen/SKILL.md (or your agent's skills folder). This skill also uses 1 other file; get the full folder from GitHub.
name
glmv-prompt-gen
description
Analyze images/videos and generate professional prompts for text-to-image and text-to-video AI tools (Midjourney, Stable Diffusion, DALL-E, Sora, Runway, Kling, Pika). Use when the user wants to generate prompts from reference images/videos, create AI art prompts, or get prompt engineering suggestions from visual content.

GLM-V Prompt Generation Skill

Analyze reference images or videos and generate professional prompts for AI image/video generation tools.

When to Use

  • Generate prompts for text-to-image tools (Midjourney, Stable Diffusion, DALL-E, etc.)
  • Generate prompts for text-to-video tools (Sora, Runway, Kling, Pika, etc.)
  • User mentions "生成prompt", "文生图prompt", "文生视频prompt", "prompt工程", "参考图生成prompt", "generate prompt"
  • User provides an image/video and wants to recreate or remix it
  • Extract prompt ideas from reference visual content

Supported Input Types

TypeFormatsMax SizeMax CountBase64
Imagejpg, png, jpeg5MB / 6000×6000px50✅
Videomp4, mkv, mov200MB—❌ (URL only)

⚠️ Images and videos cannot be used in the same request. ⚠️ Videos only support URLs — local paths and base64 are NOT supported.

📋 Output Display Rules (MANDATORY)

After running the script, you must display the full prompt output exactly as returned. Do not summarize, truncate, or only say "prompt generated". Users need the complete prompt (especially the English prompt) for direct copy/paste.

  • Show the full output: content analysis + prompt + prompt breakdown
  • In auto mode, show both text-to-image and text-to-video prompts
  • English prompts are core output and must be shown completely
  • If output was saved (-o), provide the file path and show file content

Output Modes

ModeDescription
imageGenerate prompts for text-to-image tools (default)
videoGenerate prompts for text-to-video tools
autoGenerate prompts for both image and video

Prerequisites

API Key Setup / API Key 配置(Required / 必需)

This script reads the key from the ZHIPU_API_KEY environment variable and shares it with other Zhipu skills. 脚本通过 ZHIPU_API_KEY 环境变量获取密钥,与其他智谱技能共用同一个 key。

Get Key / 获取 Key: Visit Zhipu Open Platform API Keys / 智谱开放平台 API Keys to create or copy your key.

Setup options / 配置方式(任选一种):

  1. OpenClaw config (recommended) / OpenClaw 配置(推荐): Set in openclaw.json under skills.entries.glmv-prompt-gen.env:

    json
    "glmv-prompt-gen": { "enabled": true, "env": { "ZHIPU_API_KEY": "你的密钥" } }
  2. Shell environment variable / Shell 环境变量: Add to ~/.zshrc:

    bash
    export ZHIPU_API_KEY="你的密钥"

💡 If you already configured another Zhipu skill (for example zhipu-tools or glmv-caption), they share the same ZHIPU_API_KEY, so no extra setup is needed. 💡 如果你已为其他智谱 skill(如 zhipu-tools、glmv-caption)配置过 key,它们共享同一个 ZHIPU_API_KEY,无需重复配置。

Show full SKILL.md (159 more words)Show less

How to Use

Image → Text-to-Image Prompt
bash
python scripts/prompt_gen.py --images "https://example.com/photo.jpg"
python scripts/prompt_gen.py --images /path/to/photo.png
Image → Text-to-Video Prompt
bash
python scripts/prompt_gen.py --images "https://example.com/scene.jpg" --mode video
Image → Both (Image + Video Prompts)
bash
python scripts/prompt_gen.py --images "https://example.com/photo.jpg" --mode auto
Video → Text-to-Video Prompt
bash
python scripts/prompt_gen.py --videos "https://example.com/clip.mp4" --mode video
Save Result to File
bash
python scripts/prompt_gen.py --images photo.jpg --mode image -o prompt.md
Custom Model
bash
python scripts/prompt_gen.py --images photo.jpg --model glm-4.6v-flash

Output Example (image mode)

### Content Analysis
A cyberpunk cityscape at night, with dense skyscrapers, glowing neon signs, and rain-wet streets reflecting colorful light.

### Prompt
Cyberpunk cityscape at night, towering skyscrapers with glowing neon signs,
rain-wet streets reflecting colorful lights, flying cars in the distance,
volumetric fog, dramatic lighting, ultra detailed, 8K, cinematic composition

### Prompt Breakdown
- **Subject**: Futuristic skyline with skyscrapers and neon lights
- **Style**: Cyberpunk, sci-fi
- **Color**: Cool/warm contrast with blue-purple dominance and neon accents
- **Lighting**: Neon glow, wet-surface reflections, volumetric fog
- **Composition**: Wide-angle perspective with layered depth
- **Mood**: Mysterious, futuristic, high-tech

CLI Reference

python scripts/prompt_gen.py (--images IMG [IMG...] | --videos VID [VID...]) [OPTIONS]
ParameterRequiredDescription
--images, -iOne ofImage paths or URLs (jpg/png/jpeg, base64 OK)
--videos, -vOne ofVideo URLs (mp4/mkv/mov, URL only)
--mode, -mNoOutput mode: image (default), video, or auto
--modelNoModel name (default: glm-4.6v)
--temperature, -tNoSampling temperature 0-1 (default: 0.6)
--max-tokensNoMax output tokens (default: 2048)
--thinkingNoEnable thinking/reasoning mode
--streamNoEnable streaming output
--output, -oNoSave result to file
--prettyNoPretty-print JSON error output

Error Handling

API key not configured: → Guide user to configure ZHIPU_API_KEY

Authentication failed (401/403): → API key invalid/expired → check at Zhipu API Keys / 智谱官网

Rate limit (429): → Quota exhausted → wait and retry

Content filtered: → warning field present → content blocked by safety review

Timeout: → Video processing may take time → increase timeout or use smaller files

© zai-org, Apache-2.0. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

SKILL.md and 1 other file (scripts) in skills/glmv-prompt-gen of zai-org/GLM-skills.

  • SKILL.md
  • scripts/prompt_gen.py

Open the folder on GitHubat commit 2ecd31c

Compare with similar skills

Glmv Prompt Gen next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Glmv Prompt Gen compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Glmv Prompt Gen this skillzai-org/GLM-skills476—~1.8kAutomated safety check: PassApache-2.0
Prompt EngineAgriciDaniel/claude-prompts111—~1.2kAutomated safety check: PassMIT
Nano Banana Pro Prompts Recommend SkillYouMind-OpenLab/nano-banana-pro-prompts-recommend-skill1.9k1 repos~4.1kAutomated safety check: PassNone
ImageNexus-JPF/note-companion8703 repos~3.9kAutomated safety check: PassMIT
AI Image Generationaiskillstore/marketplace4332 repos~1.1kAutomated safety check: PassNone
AI Image Prompts SkillLeoYeAI/openclaw-master-skills2.2k—~4.3kAutomated safety check: PassMIT

Similar skills

  • Prompt Engine

    AgriciDaniel/claude-prompts

    Ultimate AI prompt database and builder with 2,500+ curated prompts across 19 categories and 17 AI models (Midjourney, Flux, Leonardo AI, DALL-E, Sora, Imagen, Mystic, Stable Diffusion, Ideogram…

    111 GitHub stars~1.2k tokensUpdated 6 mo ago
    AI & LLM EngineeringAuto-check passed
  • Nano Banana Pro Prompts Recommend Skill

    YouMind-OpenLab/nano-banana-pro-prompts-recommend-skill

    Recommend suitable prompts from 10,000+ Nano Banana Pro image generation prompts based on user needs.

    1.9k GitHub starsUsed in 1 repo~4.1k tokens
    Media & CreativeAuto-check passed
  • Image

    Nexus-JPF/note-companion

    When the user wants to create, generate, edit, or optimize images for marketing — blog heroes, social graphics, product mockups, profile banners, listing visuals, or brand assets.

    870 GitHub starsUsed in 3 repos~3.9k tokens
    Media & CreativeAuto-check passed
  • AI Image Generation

    aiskillstore/marketplace

    Generate AI images with FLUX, Gemini, Grok, Seedream, Reve and 50+ models via inference.sh CLI.

    433 GitHub starsUsed in 2 repos~1.1k tokens
    Media & CreativeAuto-check passed
  • AI Image Prompts Skill

    LeoYeAI/openclaw-master-skills

    Recommend curated prompts from a 10,000+ real-world image generation prompt library.

    2.2k GitHub stars~4.3k tokensUpdated 2 mo ago
    Media & CreativeAuto-check passed
  • AI Image Generation

    aiskillstore/marketplace

    Generate AI images with GPT-Image-2, FLUX, Gemini, Grok, Seedream, Reve and 50+ models via inference.sh CLI.

    433 GitHub starsUsed in 1 repo~1.3k tokens
    Media & CreativeAuto-check passed

More from zai-org/GLM-skills

All 16 skills in this repo
  • Glmocr

    zai-org/GLM-skills

    Extract text from images using GLM-OCR API. An agent skill from zai-org/GLM-skills.

    476 GitHub stars~1.1k tokensUpdated 5 mo ago
    Auto-check passed
  • Glm Image Gen

    zai-org/GLM-skills

    Official skill for generating high-quality images from text prompts using ZhiPu GLM-Image API.

    476 GitHub stars~2.9k tokensUpdated 5 mo ago
    Auto-check passed
  • Glmocr Formula

    zai-org/GLM-skills

    Official skill for recognizing and extracting mathematical formulas from images and PDFs into LaTeX format using ZhiPu GLM-OCR API.

    476 GitHub stars~2.2k tokensUpdated 5 mo ago
    Auto-check passed
  • Glmocr Handwriting

    zai-org/GLM-skills

    Official skill for recognizing handwritten text from images using ZhiPu GLM-OCR API.

    476 GitHub stars~1.7k tokensUpdated 5 mo ago
    Auto-check passed
  • Glmocr Table

    zai-org/GLM-skills

    Official skill for recognizing and extracting tables from images and PDFs into Markdown format using ZhiPu GLM-OCR API.

    476 GitHub stars~1.7k tokensUpdated 5 mo ago
    Auto-check passed
  • Glmv Caption

    zai-org/GLM-skills

    Generate captions (descriptions) for images, videos, and documents using ZhiPu GLM-V multimodal model series.

    476 GitHub stars~2k tokensUpdated 5 mo ago
    Auto-check: notes

Questions about Glmv Prompt Gen

What does Glmv Prompt Gen do?

Analyze images/videos and generate professional prompts for text-to-image and text-to-video AI tools (Midjourney, Stable Diffusion, DALL-E, Sora, Runway, Kling, Pika). Glmv Prompt Gen is an agent skill from zai-org/GLM-skills. Analyze images/videos and generate professional prompts for text-to-image and text-to-video AI tools (Midjourney, Stable Diffusion, DALL-E, Sora, Runway, Kling, Pika).

When should I use Glmv Prompt Gen?

Glmv Prompt Gen fits situations like: the user wants to generate prompts from reference images/videos; create AI art prompts; get prompt engineering suggestions from visual content.

How do I install Glmv Prompt Gen in Claude Code?

Run `npx skills add zai-org/GLM-skills --skill glmv-prompt-gen -a claude-code`. Or copy the skill folder (skills/glmv-prompt-gen in zai-org/GLM-skills) into .claude/skills/glmv-prompt-gen in your project. Claude Code loads it when a task matches its description.

How do I install Glmv Prompt Gen in Codex?

Run `npx skills add zai-org/GLM-skills --skill glmv-prompt-gen -a codex`. Or copy the skill folder (skills/glmv-prompt-gen in zai-org/GLM-skills) into .agents/skills/glmv-prompt-gen in your project. Codex loads it when a task matches its description.

Can I use Glmv Prompt Gen in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add zai-org/GLM-skills --skill glmv-prompt-gen -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/glmv-prompt-gen, .gemini/skills/glmv-prompt-gen, .github/skills/glmv-prompt-gen and .opencode/skills/glmv-prompt-gen in your project.

What does Glmv Prompt Gen need to run?

Going by SKILL.md and its folder, Glmv Prompt Gen needs Python for the scripts in its folder, the command-line tools its instructions call (python) and credentials named ZHIPU_API_KEY. Our summary lists: Python 3; A credential in ZHIPU_API_KEY.

Does Glmv Prompt Gen access the network?

SKILL.md names 2 domains. As links in the text: bigmodel.cn and docs.bigmodel.cn. This is read from the text; nothing was executed.

Is Glmv Prompt Gen safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. The check reads SKILL.md only: the scripts in the folder are not scanned, so read them before running anything.

What licence does Glmv Prompt Gen use?

Glmv Prompt Gen is published under the Apache-2.0 licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Glmv Prompt Gen use?

About 1.8k tokens (SKILL.md is roughly 7.4k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Glmv Prompt Gen?

Skills that share tags, products or a category with Glmv Prompt Gen: Prompt Engine (AgriciDaniel/claude-prompts, 111 stars), Nano Banana Pro Prompts Recommend Skill (YouMind-OpenLab/nano-banana-pro-prompts-recommend-skill, 1.9k stars), Image (Nexus-JPF/note-companion, 870 stars) and AI Image Generation (aiskillstore/marketplace, 433 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Glmv Prompt Gen?

zai-org (a GitHub organization) maintains it in zai-org/GLM-skills, which has 476 GitHub stars. The repository holds 16 skills in this directory. The repository was last updated on April 15, 2026.

Source: zai-org/GLM-skills on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.