Agent skill

Jimeng API

by iptag in iptag/jimeng-api

Generate images using the Jimeng API based on text prompts. An agent skill from iptag/jimeng-api.

GPL-3.0Auto-check passedMedia & Creative

Install Jimeng API

skills CLI
$ npx skills add iptag/jimeng-api --skill jimeng-api -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install iptag/jimeng-api jimeng-api --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/iptag/jimeng-api.git skills-src && mkdir -p .claude/skills && cp -r skills-src/jimeng-api .claude/skills/jimeng-api && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
jimeng-api
GitHub stars
1k
Token cost
~3.3k tokens
SKILL.md length
1,121 words
Files
2 (incl. scripts)
Skills in repo
1
Repo updated
First seen
Licence
GPL-3.0

At a glance

Generate images using the Jimeng API based on text prompts. An agent skill from iptag/jimeng-api.

  • Works in 5 steps: Receive user request for image generation → Request Session ID from the user if not… → Clarify requirements → …
  • Users request AI-generated images from the Jimeng (即梦AI) service
  • SKILL.md covers Overview, When to Use This Skill, Quick Start and Parameter Usage Guidelines, plus 8 more sections
  • Runs Python scripts from its folder; calls python and pip; reaches p3-dreamina-sign.byteimg.com and p26-dreamina-sign.byteimg.com

What it does

Jimeng API is an agent skill from iptag/jimeng-api. Generate images using the Jimeng API based on text prompts. Use this skill when users request AI-generated images from the Jimeng (即梦AI) service, artwork, illustrations, or visual content creation. Supports text-to-image and image-to-image generation with customizable ratios and resolutions.

Its SKILL.md is about 3.3k tokens, which your agent loads only when the skill is triggered. The skill folder holds 2 other files, including scripts (for example `scripts/generate_image.py`).

It sits in Media & Creative, covering Image generation. The repository describes itself as: Reverse-engineered the official API for Jimeng/Dreamina’s text-to-image and image-to-image features. Drew inspiration from several experts’ projects and made some tweaks, which… The licence is GPL-3.0.

When your agent uses it

  • Users request AI-generated images from the Jimeng (即梦AI) service
  • Visual content creation

Example prompts

  • “/jimeng-api”

Requirements

  • Python 3
  • Docker

Workflow steps

5 steps, taken from the first numbered list in SKILL.md.

  1. Receive user request for image generation
  2. Request Session ID from the user if not already provided
  3. Clarify requirements
  4. Execute generation using the generate_image.py script — REMINDER: only pass parameters explicitly requested by the user; do not add/guess…
  5. Report results — show file paths only. DO NOT READ/OPEN/ANALYZE GENERATED IMAGES. DO NOT CALL ANY READ TOOL (e.g., Read, view_image). STOP…

What it can do on your machine

Read from SKILL.md and the folder at commit b9a4199. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Ships 1 file in scripts/ (Python), which the agent can run.

    Shell commands in SKILL.md call:

    • python
    • pip

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    Hosts in commands or code, which the agent is likely to contact:

    • p3-dreamina-sign.byteimg.com
    • p26-dreamina-sign.byteimg.com

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Jimeng API loads about 3.3k tokens when it runs. Until then it costs about 76 tokens; SKILL.md has 1,121 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~76
When it runs · the whole SKILL.md, loaded when a task matches
~3.3k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); the scripts in this folder are not scanned.

SKILL.md

The full file from iptag/jimeng-api at commit b9a4199, republished under its GPL-3.0 licence (© iptag). 1,121 words, ~3,273 tokens.

Download SKILL.mdSave it as .claude/skills/jimeng-api/SKILL.md (or your agent's skills folder). This skill also uses 1 other file; get the full folder from GitHub.
name
jimeng-api
description
Generate images using the Jimeng API based on text prompts. Use this skill when users request AI-generated images from the Jimeng (即梦AI) service, artwork, illustrations, or visual content creation. Supports text-to-image and image-to-image generation with customizable ratios and resolutions.
version
1.0.0
dependencies
python>=3.7, requests>=2.28.0, Pillow>=9.0.0

Jimeng API

Overview

This skill enables image generation using a locally deployed Jimeng API service (Docker). It converts text prompts into high-quality images and automatically downloads them to the project's /pic folder. The skill supports text-to-image generation, image-to-image composition, customizable aspect ratios (1:1, 16:9, etc.), and multiple resolution levels (1k, 2k, 4k).

API Endpoint: http://localhost:5100

When to Use This Skill

Use this skill when users request:

  • "使用即梦生成图片 [描述]"
  • "Generate an image using Jimeng: [description]"
  • "Create artwork showing [scene/concept]"
  • "Make an illustration of [subject]"
  • "Generate a 4K image of [description]"
  • "Transform this image to [style]" (image-to-image)
  • Any request involving Jimeng AI image generation or visual content creation

Quick Start

Prerequisites

IMPORTANT: The Jimeng API must be running locally via Docker before using this skill.

Region-specific prefixes:

  • 国内站: Direct sessionid (e.g., your_session_id)
  • 美国站: Add us- prefix (e.g., us-your_session_id)
  • 香港站: Add hk- prefix (e.g., hk-your_session_id)
  • 日本站: Add jp- prefix (e.g., jp-your_session_id)
  • 新加坡站: Add sg- prefix (e.g., sg-your_session_id)

⚠️ nanobanana Model Resolution Rules:

  • US site (us-): Fixed at 1024×1024 with 2k resolution; ignores user-provided ratio and resolution parameters
  • HK/JP/SG sites (hk-/jp-/sg-): Forced 1k resolution, but supports custom ratio parameters (e.g., 16:9, 4:3)
  • Domestic site (CN): Does not support nanobanana model; use jimeng series instead

Always ask the user for their Session ID before proceeding, as the skill does not include a pre-configured credential.

Example prompt to user:

"要使用即梦API生成图片,我需要您的Session ID。您可以从即梦网站(jimeng.jianying.com)的浏览器Cookie中获取 sessionid。

如果使用国际站,请在sessionid前添加对应前缀(us-/hk-/jp-/sg-)。

请提供您的 Session ID。"

Parameter Usage Guidelines

⚠️⚠️ IMPORTANT PARAMETER DISCIPLINE

  • ✅ ONLY PASS PARAMETERS THE USER EXPLICITLY MENTIONS.
  • ❌ DO NOT GUESS OR ADD UNSPECIFIED PARAMETERS.
  • ✅ LET THE SCRIPT USE BUILT-IN DEFAULTS when the user did not specify:
    • ratio: 1:1
    • resolution: 2k
    • model: jimeng-4.0
    • intelligent_ratio: false

Rationale: Tools may “helpfully” add options (e.g., --ratio 16:9) that the user didn’t request, overriding script defaults. This is prohibited. Pass only the parameters the user asked for; otherwise, rely on defaults.

Basic Workflow
  1. Receive user request for image generation
  2. Request Session ID from the user if not already provided
  3. Clarify requirements:
    • Text prompt (文生图) or input images (图生图)
    • Model selection (jimeng-4.0, jimeng-3.1, etc.)
    • Aspect ratio (1:1, 16:9, 4:3, etc.)
    • Resolution (1k, 2k, 4k)
    • Intelligent ratio (auto-detect based on prompt keywords)
  4. Execute generation using the generate_image.py script — REMINDER: only pass parameters explicitly requested by the user; do not add/guess any optional flags
  5. Report results — show file paths only. DO NOT READ/OPEN/ANALYZE GENERATED IMAGES. DO NOT CALL ANY READ TOOL (e.g., Read, view_image). STOP AFTER SAVING.

Image Generation Tasks

Text-to-Image Generation

Generate images from text descriptions.

Minimal default usage (no optional params):

bash
python scripts/generate_image.py text \
    "a cute cat" \
    --session-id "YOUR_SESSION_ID"

Only include optional parameters when the user explicitly requests them.

With user-specified parameters (only when requested):

bash
python scripts/generate_image.py text \
    "futuristic city at sunset with flying cars" \
    --session-id "YOUR_SESSION_ID" \
    --model "jimeng-4.0" \
    --ratio "16:9" \
    --resolution "2k"

Parameters:

  • prompt (required): Text description of the desired image
  • --session-id: Jimeng session ID (required)
  • --model: Model to use (default: jimeng-4.0)
    • Options: jimeng-5.0, jimeng-4.6, jimeng-4.5, jimeng-4.1, jimeng-4.0, jimeng-3.1, jimeng-3.0, nanobanana (international only)
  • --ratio: Aspect ratio (default: 1:1)
    • Options: 1:1, 4:3, 3:4, 16:9, 9:16, 3:2, 2:3, 21:9
  • --resolution: Resolution level (default: 2k)
    • Options: 1k, 2k, 4k
  • --intelligent-ratio: Enable smart ratio detection based on prompt keywords ⚠️ Only works for jimeng-4.0/jimeng-4.1/jimeng-4.5 models; other models will ignore this parameter
  • --negative-prompt: Negative prompt (elements to avoid)
  • --sample-strength: Sampling strength (0.0-1.0)
  • --api-url: Custom API URL (default: http://localhost:5100)
  • --output-dir: Custom output directory (defaults to project_root/pic)
Image-to-Image Composition

Transform or compose images based on text guidance.

Example user request:

"把这张照片转换成油画风格,色彩鲜艳,笔触明显"

Script usage:

bash
# Using local file
python scripts/generate_image.py image \
    "transform to oil painting style, vivid colors, visible brushstrokes" \
    --session-id "YOUR_SESSION_ID" \
    --images "/path/to/image.jpg"

# Using image URL
python scripts/generate_image.py image \
    "anime style, cute cat" \
    --session-id "YOUR_SESSION_ID" \
    --images "https://example.com/cat.jpg"

# Multiple images (up to 10)
python scripts/generate_image.py image \
    "merge these images into a cohesive scene" \
    --session-id "YOUR_SESSION_ID" \
    --images "image1.jpg" "image2.png" "image3.jpg"

Parameters:

  • Same as text-to-image, plus:
  • --images: One or more image paths or URLs (1-10 images)

Supported formats: JPG, PNG, WebP Size limit: Recommended <10MB per image

Intelligent Ratio Detection

⚠️ IMPORTANT: This feature only works with the jimeng-4.0, jimeng-4.1, and jimeng-4.5 models. Other models (jimeng-3.0, nanobanana, etc.) will ignore the --intelligent-ratio flag.

Use --intelligent-ratio to automatically select the best aspect ratio based on prompt keywords.

Example:

bash
python scripts/generate_image.py text \
    "奔跑的狮子,竖屏" \
    --session-id "YOUR_SESSION_ID" \
    --model "jimeng-4.0" \
    --intelligent-ratio
Resolution Options
ResolutionRatioDimensions
1k1:11024×1024
4:3768×1024
3:41024×768
16:91024×576
9:16576×1024
3:21024×682
2:3682×1024
21:91195×512
2k (default)1:12048×2048
16:92560×1440
4:32304×1728
4k1:14096×4096
16:95120×2880
21:96048×2592

Script Details

Location

scripts/generate_image.py

Show full SKILL.md (471 more words)Show less
Key Features
  • Automatic project root detection (looks for .git, .claude, etc.)
  • Creates /pic folder if it doesn't exist
  • Timestamps filenames to prevent overwrites (format: jimeng_YYYYMMDD_HHMMSS_N.png)
  • Automatic WebP to PNG conversion for maximum compatibility
  • Downloads all generated images from API response
  • Supports both text-to-image and image-to-image modes
  • Handles multipart/form-data for local file uploads
  • Error handling for API calls and downloads
  • Prints generation statistics
Output Format
  • Images are saved to: {project_root}/pic/jimeng_{timestamp}_{index}.png
  • All images are automatically converted to PNG format (including WebP sources)
  • Each API call generates several variations
Requirements

The script requires:

bash
pip install requests Pillow

Note: Pillow is required for WebP to PNG conversion. If not installed, WebP images will be saved as-is.

Workflow Decision Tree

User requests image generation
    ↓
Is Jimeng API running at localhost:5100?
    ├─ No → Instruct user to start Docker service
    └─ Yes → Continue
    ↓
Do we have Session ID?
    ├─ No → Request Session ID from user → Store for session
    └─ Yes → Continue
    ↓
Text-to-Image or Image-to-Image?
    ├─ Text-to-Image
    │   └─ Run: generate_image.py text "prompt" --session-id ID  (add --ratio/--resolution/--model ONLY if user explicitly requests)
    └─ Image-to-Image
        └─ Run: generate_image.py image "prompt" --session-id ID --images PATH1 [PATH2...]
    ↓
Script executes:
    1. Calls Jimeng API (文生图 or 图生图)
    2. Receives image URLs
    3. Downloads all images to /pic folder
    4. Reports file paths
    ↓
    Inform user of results
        ├─ Success → Show file paths only
        └─ Failure → Report error, suggest troubleshooting
        ↓
    HARD STOP — DO NOT READ/OPEN/ANALYZE IMAGES; DO NOT CALL `Read`/`view_image`; TASK COMPLETE

Troubleshooting

Common Issues

"Session ID required"

  • Ensure the user has provided their sessionid from 即梦/Dreamina
  • Verify correct region prefix (us-/hk-/jp-/sg- for international sites)

"Invalid session or authentication failed"

  • Session ID may have expired
  • Request user to refresh their browser and get a new sessionid
  • Verify the sessionid is copied correctly (no extra spaces)

"Error downloading image"

  • Check network connectivity
  • Verify output directory is writable
  • Image URLs may have expired (retry generation)

"Model not supported"

  • nanobanana only works with international sites (us-/hk-/jp-/sg- prefix)
  • jimeng-3.1 only works with domestic sites

"nanobanana resolution mismatch"

  • US site (us- prefix): nanobanana model only supports 1024×1024 @ 2k resolution; all ratio and resolution parameters are ignored
  • HK/JP/SG sites (hk-/jp-/sg- prefix): nanobanana model forces 1k resolution, but allows custom ratios (e.g., 16:9, 4:3)
  • If you need full control over resolution and ratio, use jimeng-4.0 model instead

"intelligent_ratio not working"

  • The --intelligent-ratio flag only works with jimeng-4.0, jimeng-4.1, and jimeng-4.5 models
  • Other models (jimeng-3.0, nanobanana, etc.) will ignore this parameter
  • Solution: Use jimeng-4.0, jimeng-4.1, or jimeng-4.5 if you need intelligent ratio detection

Best Practices

  1. Request Session ID early - Ask for it upfront if not already provided
  2. Parameter discipline: Only pass explicitly requested parameters
  3. Clarify generation mode - Determine if user wants text-to-image or image-to-image
  4. Use intelligent ratio when applicable - Enable when prompt contains orientation hints
  5. Inform about output location - Always tell users where images are saved; DO NOT read/open/analyze images; DO NOT call Read/view_image; STOP after saving
  6. Handle all variations - API returns image-urls; download and present all of them
  7. HARD STOP — REPORT FILE PATHS ONLY - DO NOT READ/OPEN/ANALYZE GENERATED IMAGES. NEVER CALL ANY READ TOOL AFTER GENERATION. TASK IS COMPLETE ONCE FILES ARE SAVED.

Example Interactions

Example 1: Simple text-to-image

User: "用即梦生成一张日落山景图"
Claude: "我来帮您使用即梦API生成图片。首先需要确认您的Session ID..."
[User provides sessionid]
Claude: [Executes generate_image.py text mode]
Claude: "✅ 成功生成4张图片!
- /project/pic/jimeng_20250116_153022_1.png
- /project/pic/jimeng_20250116_153022_2.png
- /project/pic/jimeng_20250116_153022_3.png
- /project/pic/jimeng_20250116_153022_4.png"

Example 2: High-resolution with custom ratio

User: "生成4K分辨率的21:9超宽屏壮丽山水风景"
Claude: [Executes with --resolution "4k" --ratio "21:9"]
Claude: "✅ 已生成4K超宽屏图片(6048×2592):
- 已保存到 /project/pic/jimeng_20250116_154530_1.png 等4张图片"

Example 3: Image-to-image transformation

User: "把这张猫咪照片转成动漫风格"
Claude: "请提供猫咪照片的路径或URL"
[User provides image path]
Claude: [Executes image mode with --images]
Claude: "✅ 图片风格转换完成,生成了4个动漫风格版本:
- /project/pic/jimeng_20250116_160015_1.png
- ..."

Example 4: Intelligent ratio

User: "生成一张竖屏手机壁纸,星空主题"
Claude: [Executes with --intelligent-ratio, auto-detects 9:16]
Claude: "✅ 已根据'竖屏'关键词自动选择9:16比例,生成星空壁纸4张"

API Response Format

The Jimeng API returns image variations per request:

json
{
    "created": 1763260188,
    "data": [
        {"url": "https://p3-dreamina-sign.byteimg.com/...image1.png"},
        {"url": "https://p26-dreamina-sign.byteimg.com/...image2.png"},
        {"url": "https://p26-dreamina-sign.byteimg.com/...image3.png"},
        {"url": "https://p3-dreamina-sign.byteimg.com/...image4.png"}
    ]
}

All images are automatically downloaded and saved with sequential numbering.

Security Notes

  • ⚠️ Never hardcode Session IDs in scripts or skill files
  • ⚠️ Session IDs are user-specific credentials; treat them as passwords
  • ⚠️ Ensure the local API endpoint is not exposed publicly
  • ⚠️ Image URLs from API responses may expire; download immediately

© iptag, GPL-3.0. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

SKILL.md and 1 other file (scripts) in jimeng-api of iptag/jimeng-api.

  • SKILL.md
  • scripts/generate_image.py

Open the folder on GitHubat commit b9a4199

Compare with similar skills

Jimeng API next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Jimeng API compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Jimeng API this skilliptag/jimeng-api1k—~3.3kAutomated safety check: PassGPL-3.0
Structured Image Generationbytedance/deer-flow83k5 repos~2.9kAutomated safety check: PassMIT
AI Image Generation and Editingzhayujie/CowAgent47k—~1.3kAutomated safety check: PassMIT
Canghe Comicfreestylefly/canghe-skills4618 repos~3.2kAutomated safety check: PassNone
Generate Imageynulihao/AgentSkillOS61710 repos~1.7kAutomated safety check: NotesNone
GPT Image Generation CLIwuyoscar/GPT-Image2-Skill5.7k—~2.5kAutomated safety check: NotesMIT

Similar skills

  • Structured Image Generation

    bytedance/deer-flow

    Turns an image request into a structured JSON prompt and runs a bundled Python script to generate the picture, optionally guided by reference images.

    83k GitHub starsUsed in 5 repos~2.9k tokens
    Media & CreativeAuto-check passed
  • Generates or edits images from text prompts through a Python script that picks an image backend based on which API keys are configured.

    47k GitHub stars~1.3k tokensUpdated today
    Media & CreativeAuto-check passed
  • Canghe Comic

    freestylefly/canghe-skills

    Knowledge comic creator supporting multiple art styles and tones.

    461 GitHub starsUsed in 8 repos~3.2k tokens
    Media & CreativeAuto-check passed
  • Generate Image

    ynulihao/AgentSkillOS

    Generate or edit images using AI models (FLUX, Gemini). An agent skill from ynulihao/AgentSkillOS.

    617 GitHub starsUsed in 10 repos~1.7k tokens
    Media & CreativeAuto-check: notes
  • GPT Image Generation CLI

    wuyoscar/GPT-Image2-Skill

    Generates and edits images with GPT Image 2 or 2.5 through a packaged CLI and a prompt gallery, after settling which model fits the request.

    5.7k GitHub stars~2.5k tokensUpdated 8 days ago
    Media & CreativeAuto-check: notes
  • Minimal Zine Poster Generator

    LiamGvchi/gc-minimal-zine-poster

    Creates or analyzes quiet, paper-texture zine posters with big negative space, one color accent and experimental type, returning an image prompt and the generated poster.

    7.3k GitHub stars~2.9k tokensUpdated 1 mo ago
    Media & CreativeAuto-check passed

Questions about Jimeng API

What does Jimeng API do?

Generate images using the Jimeng API based on text prompts. An agent skill from iptag/jimeng-api. Jimeng API is an agent skill from iptag/jimeng-api. Generate images using the Jimeng API based on text prompts.

When should I use Jimeng API?

Jimeng API fits situations like: users request AI-generated images from the Jimeng (即梦AI) service; visual content creation.

How do I install Jimeng API in Claude Code?

Run `npx skills add iptag/jimeng-api --skill jimeng-api -a claude-code`. Or copy the skill folder (jimeng-api in iptag/jimeng-api) into .claude/skills/jimeng-api in your project. Claude Code loads it when a task matches its description.

How do I install Jimeng API in Codex?

Run `npx skills add iptag/jimeng-api --skill jimeng-api -a codex`. Or copy the skill folder (jimeng-api in iptag/jimeng-api) into .agents/skills/jimeng-api in your project. Codex loads it when a task matches its description.

Can I use Jimeng API in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add iptag/jimeng-api --skill jimeng-api -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/jimeng-api, .gemini/skills/jimeng-api, .github/skills/jimeng-api and .opencode/skills/jimeng-api in your project.

What does Jimeng API need to run?

Going by SKILL.md and its folder, Jimeng API needs Python for the scripts in its folder and the command-line tools its instructions call (python and pip). Our summary lists: Python 3; Docker.

Does Jimeng API access the network?

SKILL.md names 2 domains. In commands or code: p3-dreamina-sign.byteimg.com and p26-dreamina-sign.byteimg.com; the agent is likely to contact these when it follows the instructions. This is read from the text; nothing was executed.

Is Jimeng API safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. The check reads SKILL.md only: the scripts in the folder are not scanned, so read them before running anything.

What licence does Jimeng API use?

Jimeng API is published under the GPL-3.0 licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Jimeng API use?

About 3.3k tokens (SKILL.md is roughly 13k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Jimeng API?

Skills that share tags, products or a category with Jimeng API: Structured Image Generation (bytedance/deer-flow, 83k stars), AI Image Generation and Editing (zhayujie/CowAgent, 47k stars), Canghe Comic (freestylefly/canghe-skills, 461 stars) and Generate Image (ynulihao/AgentSkillOS, 617 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Jimeng API?

iptag (a GitHub user) maintains it in iptag/jimeng-api, which has 1,041 GitHub stars. The repository was last updated on March 2, 2026.

Source: iptag/jimeng-api on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.