Agent skill

Qiaomu Codex Imagegen

by joeseesun in joeseesun/qiaomu-codex-imagegen

Generate or edit images with Codex's built-in image generation from any agent, through an MCP tool or a plain CLI.

MITAuto-check passedMedia & Creative

Install Qiaomu Codex Imagegen

skills CLI
$ npx skills add joeseesun/qiaomu-codex-imagegen --skill qiaomu-codex-imagegen -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install joeseesun/qiaomu-codex-imagegen qiaomu-codex-imagegen --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
qiaomu-codex-imagegen
GitHub stars
125
Token cost
~2.3k tokens
SKILL.md length
1,177 words
Files
87 (incl. scripts, references, assets)
Skills in repo
1
Repo updated
First seen
Licence
MIT

At a glance

Generate or edit images with Codex's built-in image generation from any agent, through an MCP tool or a plain CLI.

  • Works in 5 steps: Intake (no tool call yet) → Suggest at least four divergent directions → Compose the prompt → …
  • Tasks that involve Image generation
  • SKILL.md covers How to call it, The workflow: suggest, choose,…, Rules and Output contract, plus 1 more section
  • Calls codex and node

What it does

Qiaomu Codex Imagegen is an agent skill from joeseesun/qiaomu-codex-imagegen. Generate or edit images with Codex's built-in image generation from any agent, through an MCP tool or a plain CLI. For every image request it first proposes at least four divergent style directions (24 distilled templates, 48 presets, 20 Mondo poster artists), lets the user pick, expands the prompt, generates, and checks the result. Scenario craft for 视频封面, 视频海报, 小红书配图, 公众号封面, X 封面, 朋友圈海报, 书籍封面, 专辑封面, 产品海报, 人物写真, 节气海报, 展览海报, PPT 页面, 图生图. Use for 生图, 配图, 海报, 封面, "给我几个风格", "用 Codex 画一张". Excludes: editing photos…

Its SKILL.md is about 2.3k tokens, which your agent loads only when the skill is triggered. The skill folder holds 91 other files, including scripts, reference files and assets (for example `README.md` and `agents/interface.yaml`).

It sits in Media & Creative, covering Image generation and MCP servers. It works with Model Context Protocol and Xiaohongshu. The repository describes itself as: 让任何 Agent 调用 Codex 内置生图(MCP + CLI + Skill),内置小红书、视频封面、Mondo 海报技巧 · Codex image generation for any agent. The licence is MIT.

When your agent uses it

  • Tasks that involve Image generation
  • Tasks that involve MCP servers

Example prompts

  • “给我几个风格”
  • “用 Codex 画一张”
  • “/qiaomu-codex-imagegen”

Workflow steps

5 steps, taken from the step headings in SKILL.md.

  1. Intake (no tool call yet)
  2. Suggest at least four divergent directions
  3. Compose the prompt
  4. Generate
  5. Check and deliver

What it can do on your machine

Read from SKILL.md and the folder at commit c03a1cc. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Ships 1 file in scripts/, which the agent can run.

    Shell commands in SKILL.md call:

    • codex
    • node

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    Links to these hosts (documentation or services it may open):

    • github.com
    • x.com
    • vip.xiaoxiaodong.ai

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Qiaomu Codex Imagegen loads about 2.3k tokens when it runs, and up to ~77k if it reads all its reference files. Until then it costs about 168 tokens; SKILL.md has 1,177 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~168
When it runs · the whole SKILL.md, loaded when a task matches
~2.3k
With references · SKILL.md plus every file in references/, read only if the agent opens them
~77k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); the scripts in this folder are not scanned.

SKILL.md

The full file from joeseesun/qiaomu-codex-imagegen at commit c03a1cc, republished under its MIT licence (© joeseesun). 1,177 words, ~2,327 tokens.

Download SKILL.mdSave it as .claude/skills/qiaomu-codex-imagegen/SKILL.md (or your agent's skills folder). This skill also uses 86 other files; get the full folder from GitHub.
name
qiaomu-codex-imagegen
description
Generate or edit images with Codex's built-in image generation from any agent, through an MCP tool or a plain CLI. For every image request it first proposes at least four divergent style directions (24 distilled templates, 48 presets, 20 Mondo poster artists), lets the user pick, expands the prompt, generates, and checks the result. Scenario craft for 视频封面, 视频海报, 小红书配图, 公众号封面, X 封面, 朋友圈海报, 书籍封面, 专辑封面, 产品海报, 人物写真, 节气海报, 展览海报, PPT 页面, 图生图. Use for 生图, 配图, 海报, 封面, "给我几个风格", "用 Codex 画一张". Excludes: editing photos with local tools, charts or diagrams from data, UI mockups, video generation, or anything that needs an API key-based image service.
license
MIT
metadata.version
0.3.0
metadata.author
向阳乔木 (https://x.com/vista8)

Qiaomu Codex ImageGen

Codex ships an image generator that other agents cannot call. This package relays to it (codex app-server) and adds the craft that makes the pictures usable: a library of 24 templates (48 presets) distilled from 689 reference prompts, Mondo poster techniques, scenario presets, and a rule set for facts, text and identity.

How to call it

  1. MCP tools present (suggest_directions, compose_prompt, generate_image, search_prompts, get_prompt, build_prompt, list_catalog; in Claude Code mcp__qiaomu-codex-imagegen__*): use them.
  2. No MCP: run the CLI from this skill's folder: node scripts/cli.mjs suggest|compose|generate|search|prompt|list (see --help).
  3. Neither available: tell the user to install (README). Requires the Codex CLI, logged in. Never fake an image with code.

The workflow: suggest, choose, compose, generate, check

1. Intake (no tool call yet)

Extract: topic, deliverable (what it is for), ratio, text mode, facts that must be exact, references and the role of each (identity person, product real packaging, style, layout), must-keep, must-avoid.

  • Text mode: none (no text), exact_short (only the short text you list), typeset_later (text-free base with room for the title).
  • Posters, covers, cards and slides need a typographic layer. A finished poster has a hierarchy: headline, subtitle, one or two info lines, a small tag. With only a headline, or none, the result is concept art, not a poster. So for these deliverables use exact_short and give a complete short copy set: headline plus copy as lines separated by |. Take the lines from the user's facts. If they gave none, ask, or write clearly fictional sample lines and say so; never invent real prices, dates of real events, names, statistics.
  • none fits portraits and photos (T09–T11), product macros that carry no information, and base images the user will typeset. Keep every string short and read the result: models misspell, especially long Chinese.
  • Ask only blocking questions, at most three: a missing subject, a real person's identity reference, or facts that must appear. Fill low-risk gaps with a stated default.
2. Suggest at least four divergent directions

Call suggest_directions with the topic, deliverable and any headline/text mode. It returns four or more directions with different visual mechanisms (typography-led, graphic structure, photographic, product macro, material/space, plus a Mondo artist option for poster-like work). Never mix two mechanisms into one image.

Present them to the user briefly, one numbered item each:

  • name and template id,
  • one concrete sentence of what this topic would look like in that mechanism (write it yourself from mechanism and the topic, not a copy of the lock text),
  • what you still need from them (subject, data, reference photo),
  • look at one references[].image thumbnail per direction when available, so your sentence matches the real look.

Then ask which to make: one, several, "all", or "你定". Offer to exclude these and propose more (exclude).

Skip the question when the user already fixed the style ("用 Mondo 风格", "T03", "就这个风格") or said to go ahead ("直接出", "你定"). For "你定", pick the two best-fitting directions from different families and say which and why.

3. Compose the prompt

For each chosen template direction call compose_prompt with concrete variables:

  • Map the topic to one visible silhouette, action or relation, and say how it connects to the topic. At most one main symbol and one cross-over action.
  • Colours are roles (background, primary, accent, ink), not fixed hex values. Material belongs where the preset puts it; do not overlay noise on everything.
  • No invented facts: prices, statistics, titles, names, dates, weather, quotes, QR codes, logos.
  • missing_required non-empty → ask for exactly that. weak_variables → they still hold generic defaults, fill them first.
  • Preset with structural_override → rewrite the prompt by hand: delete the base sentences the override replaces, keep no conflicting structure, then send the result as raw_prompt with the 避免:… line.
  • Typographic templates with no text: tell the user the base image only satisfies the style after the title is typeset.
  • Give copy as |-separated lines; when strings include dates or numbers, list them all so the whitelist matches exactly (or pass a full text_rule).

Mondo directions skip this step: generate_image { prompt, style: <artist>, preset: <video-cover|poster|…> }.

Full rules and the canonical agent instruction: references/design-system/agent-guide.md.

Show full SKILL.md (496 more words)Show less
4. Generate

Default one image per chosen direction, up to four, each its own generate_image call (they can run in parallel). Use count only for variants of the same direction. Say that it takes 30–180 s and spends Codex quota. out_dir goes inside the user's project when they name one; reference_images take absolute paths (identity photo, product shot, style reference, or the image to edit).

Series: lock template, type hierarchy, colour roles, texture placement and cross-over action; change one or two of scene, narrative focus, hue, module span. Pass the first image as reference_images for the rest.

5. Check and deliver

Read each saved image and test it against acceptance_checks (the tool prints them): first-glance focus, relations that must hold, materials only where intended, exact text and identity, nothing invented. State pass or the specific failed relation.

With the local corpus pack installed, compare against the references: put the result next to one or two references[].image of the same template (suggest_directions lists them). If yours is plainly less designed (bare subject, no typographic hierarchy, no secondary layer), that is a failure too: complete the copy set or the missing layer and regenerate once.

On failure change only the sentence that carries that relation and regenerate once; do not stack "more premium" adjectives. Transparent regions in a result are flattened onto white automatically (transparent_background: true keeps them). Deliver absolute paths, the direction used, and the sidecar .json (exact prompt) next to each image. Only say "generated" for images you actually produced and looked at.

Rules

  • Each generation spends the user's quota: no test generations, count above 3 only if asked.
  • People: prefer silhouettes, back views, hands or symbols unless a face is asked for. Identity consistency needs an identity reference photo; without one say the person is fictional. Never make deceptive material with a real person's likeness.
  • Facts and numbers come only from the user; complex charts and long copy are typeset afterwards, not drawn by the model.
  • Reference corpus: search_prompts / get_prompt read the local pack of 689 reference prompts (installed on the user's machine only). Use them to study how a mechanism is described; never send a corpus prompt unchanged as the user's prompt.
  • Simple route: prompt + preset + style (66 styles, 20 Mondo artists) when the user wants a quick image without the direction step.

Output contract

Reply with: absolute path(s), pixel size, direction (template/preset or Mondo artist) per image, assumptions made, acceptance result, and where the prompt sidecar is. On failure give the tool's error and the one fix to try (Codex login, shorten the prompt, check the reference path).

Resources

Copyright (c) 向阳乔木. X https://x.com/vista8, GitHub https://github.com/joeseesun/. Template library distilled from public-area samples of https://vip.xiaoxiaodong.ai/open-source; Mondo material from qiaomu-mondo-poster-design (MIT).

© joeseesun, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

SKILL.md and 86 other files (scripts, references, assets) in the repository root of joeseesun/qiaomu-codex-imagegen.

  • SKILL.md
  • .gitignore
  • LICENSE
  • README.md
  • agents/interface.yaml
  • assets/qiaomu-profile/qiaomu_avatar.jpeg
  • assets/qiaomu-profile/qiaomu_reward_qr.png
  • assets/qiaomu-profile/qiaomu_wechat_public_account_qr.jpg
  • docs/samples/T01-giant-type.webp
  • docs/samples/T02-breaking-frame.webp
  • docs/samples/T03-center-light.webp
  • docs/samples/T04-ink-editorial.webp
  • docs/samples/T05-paper-terrain.webp
  • docs/samples/T06-folk-poster.webp
  • docs/samples/T07-child-drawing.webp
  • docs/samples/T08-photo-floral.webp
  • … and 71 more

Open the folder on GitHubat commit c03a1cc

Compare with similar skills

Qiaomu Codex Imagegen next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Qiaomu Codex Imagegen compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Qiaomu Codex Imagegen this skilljoeseesun/qiaomu-codex-imagegen125—~2.3kAutomated safety check: PassMIT
Image Context Runtimeshixinnt/codex-image-context-runtime100—~723Automated safety check: PassApache-2.0
Beatdesign WorkspaceBeatAPI/BeatDesign140—~1.1kAutomated safety check: PassApache-2.0
Beatdesign WorkspaceBeatAPI/BeatDesign140—~1.3kAutomated safety check: PassApache-2.0
Cullglebis/claude-skills391—~1.9kAutomated safety check: PassMIT
Gemini SkillWJZ-P/gemini-skill832—~1.1kAutomated safety check: PassMIT

Similar skills

  • Image Context Runtime

    shixinnt/codex-image-context-runtime

    A skill your agent uses for image-heavy Codex work when generated or inspected media should stay behind a bounded MCP boundary and the controlling task should receive only durable Job IDs, hashes…

    100 GitHub stars~723 tokensUpdated 1 mo ago
    Media & CreativeAuto-check passed
  • Beatdesign Workspace

    BeatAPI/BeatDesign

    Operate a local BeatDesign project through MCP when the user asks to create, organize, generate, inspect, or edit media in BeatDesign, including captions, tail-frame continuation, and opening the…

    140 GitHub stars~1.1k tokensUpdated 17 days ago
    Media & CreativeAuto-check passed
  • Beatdesign Workspace

    BeatAPI/BeatDesign

    Operate a local BeatDesign project through MCP when the user asks to create, organize, generate, inspect, or edit media in BeatDesign, including opening the exact Canvas or Editor in Codex for…

    140 GitHub stars~1.3k tokensUpdated 17 days ago
    Media & CreativeAuto-check passed
  • Cull

    glebis/claude-skills

    This skill should be used when the user wants to view, review, rate, organize, search, or export images / AI-art generations with the Cull app.

    391 GitHub stars~1.9k tokensUpdated 2 days ago
    Media & CreativeAuto-check passed
  • Gemini Skill

    WJZ-P/gemini-skill

    通过 Gemini 官网(gemini.google.com)执行生图、对话等操作。用户提到"生图/画图/绘图/nano banana/nanobanana/生成图片"等关键词时触发。操作方式分三级优先级:首选 MCP 工具 → 次选 Skill 脚本 → 最次连接 Skill 浏览器手动操作(需用户授权)。禁止自行启动外部浏览器访问 Gemini。

    832 GitHub stars~1.1k tokensUpdated 22 days ago
    Agent WorkflowsAuto-check passed
  • Comfy

    Comfy-Org/comfy-skills

    Generate images, video, audio, and 3D with Comfy Cloud — search hundreds of models and workflow templates, run custom ComfyUI workflows, and manage generation jobs through the hosted Comfy Cloud MCP…

    222 GitHub stars~1.4k tokensUpdated 2 days ago
    Agent WorkflowsAuto-check passed

Questions about Qiaomu Codex Imagegen

What does Qiaomu Codex Imagegen do?

Generate or edit images with Codex's built-in image generation from any agent, through an MCP tool or a plain CLI. Qiaomu Codex Imagegen is an agent skill from joeseesun/qiaomu-codex-imagegen. Generate or edit images with Codex's built-in image generation from any agent, through an MCP tool or a plain CLI.

When should I use Qiaomu Codex Imagegen?

Qiaomu Codex Imagegen fits situations like: tasks that involve Image generation; tasks that involve MCP servers.

How do I install Qiaomu Codex Imagegen in Claude Code?

Run `npx skills add joeseesun/qiaomu-codex-imagegen --skill qiaomu-codex-imagegen -a claude-code`. Or copy the skill folder (the joeseesun/qiaomu-codex-imagegen repository) into .claude/skills/qiaomu-codex-imagegen in your project. Claude Code loads it when a task matches its description.

How do I install Qiaomu Codex Imagegen in Codex?

Run `npx skills add joeseesun/qiaomu-codex-imagegen --skill qiaomu-codex-imagegen -a codex`. Or copy the skill folder (the joeseesun/qiaomu-codex-imagegen repository) into .agents/skills/qiaomu-codex-imagegen in your project. Codex loads it when a task matches its description.

Can I use Qiaomu Codex Imagegen in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add joeseesun/qiaomu-codex-imagegen --skill qiaomu-codex-imagegen -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/qiaomu-codex-imagegen, .gemini/skills/qiaomu-codex-imagegen, .github/skills/qiaomu-codex-imagegen and .opencode/skills/qiaomu-codex-imagegen in your project.

What does Qiaomu Codex Imagegen need to run?

Going by SKILL.md and its folder, Qiaomu Codex Imagegen needs the command-line tools its instructions call (codex and node).

Does Qiaomu Codex Imagegen access the network?

SKILL.md names 3 domains. As links in the text: github.com, x.com and vip.xiaoxiaodong.ai. This is read from the text; nothing was executed.

Is Qiaomu Codex Imagegen safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. The check reads SKILL.md only: the scripts in the folder are not scanned, so read them before running anything.

What licence does Qiaomu Codex Imagegen use?

Qiaomu Codex Imagegen is published under the MIT licence (declared in SKILL.md). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Qiaomu Codex Imagegen use?

About 2.3k tokens (SKILL.md is roughly 9.3k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full. Its references folder adds about 75k tokens, read only when the agent opens those files.

What are the alternatives to Qiaomu Codex Imagegen?

Skills that share tags, products or a category with Qiaomu Codex Imagegen: Image Context Runtime (shixinnt/codex-image-context-runtime, 100 stars), Beatdesign Workspace (BeatAPI/BeatDesign, 140 stars), Beatdesign Workspace (BeatAPI/BeatDesign, 140 stars) and Cull (glebis/claude-skills, 391 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Qiaomu Codex Imagegen?

joeseesun (a GitHub user) maintains it in joeseesun/qiaomu-codex-imagegen, which has 125 GitHub stars. The repository was last updated on October 4, 2026.

Source: joeseesun/qiaomu-codex-imagegen on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.