Agent skill

Gpt Image 1 5

by intellectronica in intellectronica/agent-skills

Generate and edit images using OpenAI's GPT Image 1.5 model.

CC0-1.0Auto-check passedMedia & Creative

Install Gpt Image 1 5

skills CLI
$ npx skills add intellectronica/agent-skills --skill gpt-image-1-5 -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install intellectronica/agent-skills gpt-image-1-5 --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/intellectronica/agent-skills.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/gpt-image-1-5 .claude/skills/gpt-image-1-5 && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
gpt-image-1-5
GitHub stars
295
Token cost
~1.5k tokens
SKILL.md length
527 words
Files
2 (incl. scripts)
Skills in repo
21
Repo updated
First seen
Licence
CC0-1.0

At a glance

Generate and edit images using OpenAI's GPT Image 1.5 model.

  • Works in 2 steps: api-key argument (use if user provided… → OPENAI_API_KEY environment variable
  • The user asks to generate
  • SKILL.md covers Usage, Parameters, API Key and Filename Generation, plus 4 more sections
  • Runs Python scripts from its folder; calls uv; needs OPENAI_API_KEY

What it does

Gpt Image 1 5 is an agent skill from intellectronica/agent-skills. Generate and edit images using OpenAI's GPT Image 1.5 model. Use when the user asks to generate, create, edit, modify, change, alter, or update images. Also use when user references an existing image file and asks to modify it in any way (e.g., "modify this image", "change the background", "replace X with Y"). Supports text-to-image generation and image editing with optional mask. DO NOT read the image file first - use this skill directly with the --input-image parameter.

Its SKILL.md is about 1.5k tokens, which your agent loads only when the skill is triggered. The skill folder holds 2 other files, including scripts (for example `scripts/generate_image.py`).

It sits in Media & Creative, covering Image generation and Image editing. It works with OpenAI. The repository describes itself as: @intellectronica's agent skills. The licence is CC0-1.0.

When your agent uses it

  • The user asks to generate
  • User references an existing image file and asks to modify it in any way (e.g.
  • Modify this image
  • Change the background

Example prompts

  • “modify this image”
  • “change the background”
  • “replace X with Y”
  • “/gpt-image-1-5”

Requirements

  • Python 3
  • A credential in OPENAI_API_KEY

Workflow steps

2 steps, taken from the first numbered list in SKILL.md.

  1. api-key argument (use if user provided key in chat)
  2. OPENAI_API_KEY environment variable

What it can do on your machine

Read from SKILL.md and the folder at commit 9b0e00a. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Ships 1 file in scripts/ (Python), which the agent can run.

    Shell commands in SKILL.md call:

    • uv

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md. Its commands use uv, which can reach the network depending on how they are called.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names these keys or tokens, usually read from environment variables:

    • OPENAI_API_KEY

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Gpt Image 1 5 loads about 1.5k tokens when it runs. Until then it costs about 123 tokens; SKILL.md has 527 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~123
When it runs · the whole SKILL.md, loaded when a task matches
~1.5k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); the scripts in this folder are not scanned.

SKILL.md

The full file from intellectronica/agent-skills at commit 9b0e00a, republished under its CC0-1.0 licence (© intellectronica). 527 words, ~1,549 tokens.

Download SKILL.mdSave it as .claude/skills/gpt-image-1-5/SKILL.md (or your agent's skills folder). This skill also uses 1 other file; get the full folder from GitHub.
name
gpt-image-1-5
description
Generate and edit images using OpenAI's GPT Image 1.5 model. Use when the user asks to generate, create, edit, modify, change, alter, or update images. Also use when user references an existing image file and asks to modify it in any way (e.g., "modify this image", "change the background", "replace X with Y"). Supports text-to-image generation and image editing with optional mask. DO NOT read the image file first - use this skill directly with the --input-image parameter.

GPT Image 1.5 - Image Generation & Editing

Generate new images or edit existing ones using OpenAI's GPT Image 1.5 model.

  • Generation: Uses the Responses API with image_generation tool
  • Editing: Uses the Image API for reliable mask-based inpainting

Usage

Run the script using absolute path (do NOT cd to skill directory first):

Generate new image:

bash
uv run ~/.claude/skills/gpt-image-1-5/scripts/generate_image.py --prompt "your image description" --filename "output-name.png" [--quality low|medium|high] [--size 1024x1024|1024x1536|1536x1024|auto] [--background transparent|opaque|auto] [--api-key KEY]

Edit existing image (without mask - full image edit):

bash
uv run ~/.claude/skills/gpt-image-1-5/scripts/generate_image.py --prompt "editing instructions" --filename "output-name.png" --input-image "path/to/input.png" [--size 1024x1024|1024x1536|1536x1024|auto] [--api-key KEY]

Edit existing image (with mask - precise inpainting):

bash
uv run ~/.claude/skills/gpt-image-1-5/scripts/generate_image.py --prompt "what to put in masked area" --filename "output-name.png" --input-image "path/to/input.png" --mask "path/to/mask.png" [--size 1024x1024|1024x1536|1536x1024|auto] [--api-key KEY]

Important: Always run from the user's current working directory so images are saved where the user is working, not in the skill directory.

Parameters

Quality Options
  • low - Fastest generation, lower quality
  • medium (default) - Balanced quality and speed
  • high - Best quality, slower generation

Map user requests:

  • No mention of quality -> medium
  • "quick", "fast", "draft" -> low
  • "high quality", "best", "detailed", "high-res" -> high
Size Options
  • 1024x1024 (default) - Square format
  • 1024x1536 - Portrait format
  • 1536x1024 - Landscape format
  • auto - Let the model decide based on prompt

Map user requests:

  • No mention of size -> 1024x1024
  • "square" -> 1024x1024
  • "portrait", "vertical", "tall" -> 1024x1536
  • "landscape", "horizontal", "wide" -> 1536x1024
Background Options (generation only)
  • auto (default) - Model decides
  • transparent - Transparent background (PNG/WebP output)
  • opaque - Solid background

API Key

The script checks for API key in this order:

  1. --api-key argument (use if user provided key in chat)
  2. OPENAI_API_KEY environment variable

If neither is available, the script exits with an error message.

Filename Generation

Generate filenames with the pattern: yyyy-mm-dd-hh-mm-ss-name.png

Format: {timestamp}-{descriptive-name}.png

  • Timestamp: Current date/time in format yyyy-mm-dd-hh-mm-ss (24-hour format)
  • Name: Descriptive lowercase text with hyphens
  • Keep the descriptive part concise (1-5 words typically)
  • Use context from user's prompt or conversation
  • If unclear, use random identifier (e.g., x9k2, a7b3)

Examples:

  • Prompt "A serene Japanese garden" -> 2025-12-17-14-23-05-japanese-garden.png
  • Prompt "sunset over mountains" -> 2025-12-17-15-30-12-sunset-mountains.png
  • Prompt "create an image of a robot" -> 2025-12-17-16-45-33-robot.png
  • Unclear context -> 2025-12-17-17-12-48-x9k2.png

Image Editing

Both editing modes use the Image API (images.edit endpoint) with gpt-image-1.5 for reliable results.

Show full SKILL.md (223 more words)Show less
Without Mask (Full Image Edit)

When the user wants to modify an existing image without specifying exact regions:

  1. Use --input-image parameter with the path to the image
  2. The prompt should contain editing instructions (e.g., "make the sky more dramatic", "change to cartoon style")
  3. A fully transparent mask is auto-generated, allowing the model to edit the entire image
With Mask (Precise Inpainting)

When the user wants to edit specific regions:

  1. Use --input-image parameter with the path to the image
  2. Use --mask parameter with a PNG mask file
  3. The mask should have transparent areas (alpha=0) where edits should occur
  4. The prompt describes what should appear in the masked region

Common editing tasks: add/remove elements, change style, adjust colors, replace backgrounds, etc.

Prompt Handling

For generation: Pass user's image description as-is to --prompt. Only rework if clearly insufficient.

For editing: Pass editing instructions in --prompt (e.g., "add a rainbow in the sky", "make it look like a watercolor painting")

Preserve user's creative intent in both cases.

Output

  • Saves PNG to current directory (or specified path if filename includes directory)
  • Script outputs the full path to the generated image
  • Do not read the image back - just inform the user of the saved path

Examples

Generate new image:

bash
uv run ~/.claude/skills/gpt-image-1-5/scripts/generate_image.py --prompt "A serene Japanese garden with cherry blossoms" --filename "2025-12-17-14-23-05-japanese-garden.png" --quality high --size 1536x1024

Generate with transparent background:

bash
uv run ~/.claude/skills/gpt-image-1-5/scripts/generate_image.py --prompt "A cute cartoon cat mascot" --filename "2025-12-17-14-25-30-cat-mascot.png" --background transparent --quality high

Edit existing image (full image):

bash
uv run ~/.claude/skills/gpt-image-1-5/scripts/generate_image.py --prompt "make the sky more dramatic with storm clouds" --filename "2025-12-17-14-27-00-dramatic-sky.png" --input-image "original-photo.jpg"

Edit with mask (inpainting):

bash
uv run ~/.claude/skills/gpt-image-1-5/scripts/generate_image.py --prompt "a flamingo swimming" --filename "2025-12-17-14-30-00-lounge-flamingo.png" --input-image "lounge.png" --mask "mask.png"

© intellectronica, CC0-1.0. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

SKILL.md and 1 other file (scripts) in skills/gpt-image-1-5 of intellectronica/agent-skills.

  • SKILL.md
  • scripts/generate_image.py

Open the folder on GitHubat commit 9b0e00a

Compare with similar skills

Gpt Image 1 5 next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Gpt Image 1 5 compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Gpt Image 1 5 this skillintellectronica/agent-skills295—~1.5kAutomated safety check: PassCC0-1.0
AI Image Generation and Editingzhayujie/CowAgent47k—~1.3kAutomated safety check: PassMIT
GPT Image Generation CLIwuyoscar/GPT-Image2-Skill5.7k—~2.5kAutomated safety check: NotesMIT
BlockRun Image GenerationBlockRunAI/ClawRouter6.6k—~2.1kAutomated safety check: PassMIT
Codex API Image Generatoryc-duan/api-image101—~4.9kAutomated safety check: PassMIT
ImagegenJetBrains/skills3663 repos~2.5kAutomated safety check: PassApache-2.0

Similar skills

  • Generates or edits images from text prompts through a Python script that picks an image backend based on which API keys are configured.

    47k GitHub stars~1.3k tokensUpdated yesterday
    Media & CreativeAuto-check passed
  • GPT Image Generation CLI

    wuyoscar/GPT-Image2-Skill

    Generates and edits images with GPT Image 2 or 2.5 through a packaged CLI and a prompt gallery, after settling which model fits the request.

    5.7k GitHub stars~2.5k tokensUpdated 10 days ago
    Media & CreativeAuto-check: notes
  • BlockRun Image Generation

    BlockRunAI/ClawRouter

    Generates or edits images through ClawRouter's local image API, with a choice of models and sizes and payment handled automatically through x402.

    6.6k GitHub stars~2.1k tokensUpdated 5 days ago
    Media & CreativeAuto-check passed
  • Codex API Image Generator

    yc-duan/api-image

    Replaces Codex's built-in image tool with provider-based generation and editing, fetching reference images first whenever visual accuracy actually matters.

    101 GitHub stars~4.9k tokensUpdated 5 mo ago
    Media & CreativeAuto-check passed
  • Imagegen

    JetBrains/skills

    Official

    A skill your agent uses when the user asks to generate or edit images via the OpenAI Image API (for example: generate image, edit/inpaint/mask, background removal or replacement, transparent…

    366 GitHub starsUsed in 3 repos~2.5k tokens
    Media & CreativeAuto-check passed
  • BananaTape Image Editor CLI

    NomaDamas/bananatape

    Drives the BananaTape CLI to create, launch, list and delete local AI image editing projects from an agent, including headless smoke tests.

    188 GitHub stars~547 tokensUpdated 2 mo ago
    Media & CreativeAuto-check passed

More from intellectronica/agent-skills

All 21 skills in this repo
  • Beautiful Mermaid

    intellectronica/agent-skills

    Render Mermaid diagrams as SVG and PNG using the Beautiful Mermaid library.

    295 GitHub starsUsed in 2 repos~1.3k tokens
    Auto-check passed
  • Nano Banana 2

    intellectronica/agent-skills

    Generate and edit images using Google's Nano Banana 2 (Gemini 3.1 Flash Image Preview) API.

    295 GitHub stars~972 tokensUpdated 5 mo ago
    Auto-check passed
  • Nano Banana Pro

    intellectronica/agent-skills

    Generate and edit images using Google's Nano Banana Pro (Gemini 3 Pro Image) API.

    295 GitHub stars~1.1k tokensUpdated 5 mo ago
    Auto-check passed
  • Notion API

    intellectronica/agent-skills

    This skill provides comprehensive instructions for interacting with the Notion API via REST calls.

    295 GitHub starsUsed in 1 repo~3.7k tokens
    Auto-check passed
  • Markdown Converter

    intellectronica/agent-skills

    Convert documents and files to Markdown using markitdown. An agent skill from intellectronica/agent-skills.

    295 GitHub starsUsed in 4 repos~492 tokens
    Auto-check passed
  • Youtube Transcript

    intellectronica/agent-skills

    Extract transcripts from YouTube videos. An agent skill from intellectronica/agent-skills.

    295 GitHub starsUsed in 2 repos~394 tokens
    Auto-check passed

Works with

Questions about Gpt Image 1 5

What does Gpt Image 1 5 do?

Generate and edit images using OpenAI's GPT Image 1.5 model. Gpt Image 1 5 is an agent skill from intellectronica/agent-skills.5 model.

When should I use Gpt Image 1 5?

Gpt Image 1 5 fits situations like: the user asks to generate; user references an existing image file and asks to modify it in any way (e.g; modify this image; change the background.

How do I install Gpt Image 1 5 in Claude Code?

Run `npx skills add intellectronica/agent-skills --skill gpt-image-1-5 -a claude-code`. Or copy the skill folder (skills/gpt-image-1-5 in intellectronica/agent-skills) into .claude/skills/gpt-image-1-5 in your project. Claude Code loads it when a task matches its description.

How do I install Gpt Image 1 5 in Codex?

Run `npx skills add intellectronica/agent-skills --skill gpt-image-1-5 -a codex`. Or copy the skill folder (skills/gpt-image-1-5 in intellectronica/agent-skills) into .agents/skills/gpt-image-1-5 in your project. Codex loads it when a task matches its description.

Can I use Gpt Image 1 5 in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add intellectronica/agent-skills --skill gpt-image-1-5 -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/gpt-image-1-5, .gemini/skills/gpt-image-1-5, .github/skills/gpt-image-1-5 and .opencode/skills/gpt-image-1-5 in your project.

What does Gpt Image 1 5 need to run?

Going by SKILL.md and its folder, Gpt Image 1 5 needs Python for the scripts in its folder, the command-line tools its instructions call (uv) and credentials named OPENAI_API_KEY. Our summary lists: Python 3; A credential in OPENAI_API_KEY.

Does Gpt Image 1 5 access the network?

SKILL.md contains no URLs. Its commands use uv, which can reach the network depending on how they are called. This is read from the text; nothing was executed.

Is Gpt Image 1 5 safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. The check reads SKILL.md only: the scripts in the folder are not scanned, so read them before running anything.

What licence does Gpt Image 1 5 use?

Gpt Image 1 5 is published under the CC0-1.0 licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Gpt Image 1 5 use?

About 1.5k tokens (SKILL.md is roughly 6.2k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Gpt Image 1 5?

Skills that share tags, products or a category with Gpt Image 1 5: AI Image Generation and Editing (zhayujie/CowAgent, 47k stars), GPT Image Generation CLI (wuyoscar/GPT-Image2-Skill, 5.7k stars), BlockRun Image Generation (BlockRunAI/ClawRouter, 6.6k stars) and Codex API Image Generator (yc-duan/api-image, 101 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Gpt Image 1 5?

intellectronica (a GitHub user) maintains it in intellectronica/agent-skills, which has 295 GitHub stars. The repository holds 21 skills in this directory. The repository was last updated on April 25, 2026.

Source: intellectronica/agent-skills on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.