Agent skill

Openai Image Gen

by swarmclawai in swarmclawai/swarmclaw

Generate images via OpenAI Images API (GPT Image, DALL-E 3, DALL-E 2).

MITAuto-check passedMedia & Creative

Install Openai Image Gen

skills CLI
$ npx skills add swarmclawai/swarmclaw --skill openai-image-gen -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install swarmclawai/swarmclaw openai-image-gen --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/swarmclawai/swarmclaw.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/openai-image-gen .claude/skills/openai-image-gen && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
openai-image-gen
GitHub stars
689
Token cost
~705 tokens
SKILL.md length
150 words
Files
2 (incl. scripts)
Skills in repo
12
Repo updated
First seen
Licence
MIT

At a glance

Generate images via OpenAI Images API (GPT Image, DALL-E 3, DALL-E 2).

  • Asked to generate images with OpenAI and an OPENAIAPIKEY is available
  • SKILL.md covers Run, Useful Flags, Model-Specific Parameters and Output
  • Runs Python scripts from its folder; calls python3
  • Tasks that involve Image generation

What it does

Openai Image Gen is an agent skill from swarmclawai/swarmclaw. Generate images via OpenAI Images API (GPT Image, DALL-E 3, DALL-E 2). Supports batch generation with random prompt sampler and HTML gallery output. Use when asked to generate images with OpenAI and an OPENAIAPIKEY is available.

Its SKILL.md is about 710 tokens, which your agent loads only when the skill is triggered. The skill folder holds 2 other files, including scripts (for example `scripts/gen.py`).

It sits in Media & Creative, covering Image generation. It works with OpenAI and Model Context Protocol. The repository describes itself as: Open-source self-hosted AI agent runtime and multi-agent framework for autonomous agent swarms. Agent memory, MCP tools, schedules, delegation, and 23+ LLM providers (Claude… The licence is MIT.

When your agent uses it

  • Asked to generate images with OpenAI and an OPENAIAPIKEY is available
  • Tasks that involve Image generation

Example prompts

  • “/openai-image-gen”

Requirements

  • Python 3
  • A credential in OPENAI_API_KEY

What it can do on your machine

Read from SKILL.md and the folder at commit ed38ba5. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Ships 1 file in scripts/ (Python), which the agent can run.

    Shell commands in SKILL.md call:

    • python3

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Openai Image Gen loads about 705 tokens when it runs. Until then it costs about 62 tokens; SKILL.md has 150 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~62
When it runs · the whole SKILL.md, loaded when a task matches
~705

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); the scripts in this folder are not scanned.

SKILL.md

The full file from swarmclawai/swarmclaw at commit ed38ba5, republished under its MIT licence (© swarmclawai). 150 words, ~705 tokens.

Download SKILL.mdSave it as .claude/skills/openai-image-gen/SKILL.md (or your agent's skills folder). This skill also uses 1 other file; get the full folder from GitHub.
name
openai-image-gen
description
Generate images via OpenAI Images API (GPT Image, DALL-E 3, DALL-E 2). Supports batch generation with random prompt sampler and HTML gallery output. Use when asked to generate images with OpenAI and an OPENAI_API_KEY is available.

OpenAI Image Gen

Generate images via the OpenAI Images API with an HTML gallery viewer.

Run

Note: Image generation can take longer than typical timeouts. Set a higher timeout when running via shell (e.g., 300 seconds).

bash
python3 {baseDir}/scripts/gen.py

Useful Flags

bash
# GPT image models with various options
python3 {baseDir}/scripts/gen.py --count 16 --model gpt-image-1
python3 {baseDir}/scripts/gen.py --prompt "ultra-detailed studio photo of a lobster astronaut" --count 4
python3 {baseDir}/scripts/gen.py --size 1536x1024 --quality high --out-dir ./out/images
python3 {baseDir}/scripts/gen.py --model gpt-image-1.5 --background transparent --output-format webp

# DALL-E 3 (note: count is automatically limited to 1)
python3 {baseDir}/scripts/gen.py --model dall-e-3 --quality hd --size 1792x1024 --style vivid
python3 {baseDir}/scripts/gen.py --model dall-e-3 --style natural --prompt "serene mountain landscape"

# DALL-E 2
python3 {baseDir}/scripts/gen.py --model dall-e-2 --size 512x512 --count 4

Model-Specific Parameters

Size
  • GPT image models (gpt-image-1, gpt-image-1-mini, gpt-image-1.5): 1024x1024, 1536x1024 (landscape), 1024x1536 (portrait), or auto. Default: 1024x1024
  • dall-e-3: 1024x1024, 1792x1024, or 1024x1792. Default: 1024x1024
  • dall-e-2: 256x256, 512x512, or 1024x1024. Default: 1024x1024
Quality
  • GPT image models: auto, high, medium, or low. Default: high
  • dall-e-3: hd or standard. Default: standard
  • dall-e-2: standard only
Other Parameters
  • GPT image models support --background (transparent, opaque, auto) and --output-format (png, jpeg, webp)
  • dall-e-3 supports --style (vivid for hyper-real, natural for more natural looking)
  • dall-e-3 only supports n=1; the script automatically limits count to 1

Output

  • Image files (*.png, *.jpeg, or *.webp depending on model and format)
  • prompts.json (prompt-to-file mapping)
  • index.html (thumbnail gallery — open in browser to review)

© swarmclawai, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

SKILL.md and 1 other file (scripts) in skills/openai-image-gen of swarmclawai/swarmclaw.

  • SKILL.md
  • scripts/gen.py

Open the folder on GitHubat commit ed38ba5

Compare with similar skills

Openai Image Gen next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Openai Image Gen compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Openai Image Gen this skillswarmclawai/swarmclaw689—~705Automated safety check: PassMIT
Higgsfield Gpt Image 2OSideMedia/higgsfield-ai-prompt-skill713—~6.5kAutomated safety check: PassMIT
Scenario Gpt Imagescenario-labs/skills946—~2.7kAutomated safety check: PassMIT
Ag2 Use Builtin Toolsag2ai/build-with-ag2252—~1.3kAutomated safety check: PassApache-2.0
Character Refseternityspring/shuohao-skills4.3k—~1.7kAutomated safety check: WarnApache-2.0
Local AI Useamd/skills408—~5kAutomated safety check: NotesMIT

Similar skills

  • Higgsfield Gpt Image 2

    OSideMedia/higgsfield-ai-prompt-skill

    A skill your agent uses when the user mentions GPT Image 2.0 or GPT Image 2.5, gpt-image-2, gpt-image-2.5, gptimage25, GPT-Image-2 prompts, the Flare / Sunburst variants, a transparent-background…

    713 GitHub stars~6.5k tokensUpdated 14 days ago
    Media & CreativeAuto-check passed
  • Scenario Gpt Image

    scenario-labs/skills

    A skill your agent uses when generating or editing images with OpenAI's GPT Image models on Scenario via MCP: text-to-image, edits from reference images, inpainting with an alpha mask, in-image text…

    946 GitHub stars~2.7k tokensUpdated yesterday
    Media & CreativeAuto-check passed
  • Ag2 Use Builtin Tools

    ag2ai/build-with-ag2

    Wire AG2 beta's shipped tools into an Agent — both provider-native server-side tools (web search, web fetch, code execution, MCP, image generation, memory) and locally-executed common toolkits…

    252 GitHub stars~1.3k tokensUpdated 1 mo ago
    Productivity & AutomationAuto-check passed
  • Character Refs

    eternityspring/shuohao-skills

    给任何故事里的角色真出参考图(小说改编、自己原创的故事、单独设计一个角色都行,不需要小说原文): 一段话描述角色,拆成分层字段、补全后确认, 先出一张正面全身锚点,其余视图(大头照、90° 侧面、背面、细节、45° 大头照)都只参考这张锚点, 按需分档出图。每张图带标识、可单独重出,重出后自动标出哪些图过期。

    4.3k GitHub stars~1.7k tokensUpdated 2 days ago
    Media & CreativeAuto-check: warnings
  • Local AI Use

    amd/skills

    Makes this agent generate images, transcribe audio, and synthesize speech on the user's own machine through a local Lemonade Server instead of a paid cloud API.

    408 GitHub stars~5k tokensUpdated yesterday
    Media & CreativeAuto-check: notes
  • Image Gen

    open-octo/octo-agent

    Acquire images as files — generate them with an AI image model (14 providers: OpenAI/gpt-image, Gemini, Qwen, Zhipu, Volcengine, Stability, FLUX, Ideogram, MiniMax, and more), search openly-licensed…

    125 GitHub stars~3.1k tokensUpdated yesterday
    Media & CreativeAuto-check: notes

More from swarmclawai/swarmclaw

All 12 skills in this repo
  • Skill Creator

    swarmclawai/swarmclaw

    Create, edit, improve, or audit skills for SwarmClaw agents.

    689 GitHub stars~1.4k tokensUpdated 3 mo ago
    Auto-check passed
  • Nano Banana Pro

    swarmclawai/swarmclaw

    Generate or edit images via Gemini 3 Pro Image (Nano Banana Pro).

    689 GitHub stars~481 tokensUpdated 3 mo ago
    Auto-check passed
  • Coding Agent

    swarmclawai/swarmclaw

    Delegate coding tasks to external coding agents (Claude Code, Codex, Pi, OpenCode) via shell.

    689 GitHub stars~900 tokensUpdated 3 mo ago
    Auto-check passed
  • GitHub

    swarmclawai/swarmclaw

    GitHub operations via gh CLI: issues, PRs, CI runs, code review, API queries.

    689 GitHub stars~820 tokensUpdated 3 mo ago
    Auto-check passed
  • Swarmclaw

    swarmclawai/swarmclaw

    AI agent runtime and multi-agent orchestration platform. An agent skill from swarmclawai/swarmclaw.

    689 GitHub stars~2k tokensUpdated 3 mo ago
    Auto-check passed
  • Swarmclaw

    swarmclawai/swarmclaw

    Manage your SwarmClaw agent fleet — agents, tasks, chats, chatrooms, goals, schedules, memory, wallets, connectors, autonomy, and 40+ more command groups.

    689 GitHub stars~4.2k tokensUpdated 3 mo ago
    Auto-check passed

Questions about Openai Image Gen

What does Openai Image Gen do?

Generate images via OpenAI Images API (GPT Image, DALL-E 3, DALL-E 2). Openai Image Gen is an agent skill from swarmclawai/swarmclaw. Generate images via OpenAI Images API (GPT Image, DALL-E 3, DALL-E 2).

When should I use Openai Image Gen?

Openai Image Gen fits situations like: asked to generate images with OpenAI and an OPENAIAPIKEY is available; tasks that involve Image generation.

How do I install Openai Image Gen in Claude Code?

Run `npx skills add swarmclawai/swarmclaw --skill openai-image-gen -a claude-code`. Or copy the skill folder (skills/openai-image-gen in swarmclawai/swarmclaw) into .claude/skills/openai-image-gen in your project. Claude Code loads it when a task matches its description.

How do I install Openai Image Gen in Codex?

Run `npx skills add swarmclawai/swarmclaw --skill openai-image-gen -a codex`. Or copy the skill folder (skills/openai-image-gen in swarmclawai/swarmclaw) into .agents/skills/openai-image-gen in your project. Codex loads it when a task matches its description.

Can I use Openai Image Gen in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add swarmclawai/swarmclaw --skill openai-image-gen -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/openai-image-gen, .gemini/skills/openai-image-gen, .github/skills/openai-image-gen and .opencode/skills/openai-image-gen in your project.

What does Openai Image Gen need to run?

Going by SKILL.md and its folder, Openai Image Gen needs Python for the scripts in its folder and the command-line tools its instructions call (python3). Our summary lists: Python 3; A credential in OPENAI_API_KEY.

Does Openai Image Gen access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is Openai Image Gen safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. The check reads SKILL.md only: the scripts in the folder are not scanned, so read them before running anything.

What licence does Openai Image Gen use?

Openai Image Gen is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Openai Image Gen use?

About 705 tokens (SKILL.md is roughly 2.8k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Openai Image Gen?

Skills that share tags, products or a category with Openai Image Gen: Higgsfield Gpt Image 2 (OSideMedia/higgsfield-ai-prompt-skill, 713 stars), Scenario Gpt Image (scenario-labs/skills, 946 stars), Ag2 Use Builtin Tools (ag2ai/build-with-ag2, 252 stars) and Character Refs (eternityspring/shuohao-skills, 4.3k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Openai Image Gen?

swarmclawai (a GitHub organization) maintains it in swarmclawai/swarmclaw, which has 689 GitHub stars. The repository holds 12 skills in this directory. The repository was last updated on June 30, 2026.

Source: swarmclawai/swarmclaw on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.