Agent skill

Image Gen

by nicknisi in nicknisi/claude-plugins

Generate or edit images via Google Gemini (nano-banana-pro) or OpenAI gpt-image-2.

MITAuto-check passedMedia & Creative

Install Image Gen

skills CLI
$ npx skills add nicknisi/claude-plugins --skill image-gen -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install nicknisi/claude-plugins image-gen --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/nicknisi/claude-plugins.git skills-src && mkdir -p .claude/skills && cp -r skills-src/plugins/image-gen/skills/image-gen .claude/skills/image-gen && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
image-gen
GitHub stars
114
Token cost
~502 tokens
SKILL.md length
166 words
Files
4 (incl. scripts)
Skills in repo
12
Repo updated
First seen
Licence
MIT

At a glance

Generate or edit images via Google Gemini (nano-banana-pro) or OpenAI gpt-image-2.

  • Make illustration
  • SKILL.md covers Which provider to pick, Usage, API keys and Conventions
  • Runs JavaScript and TypeScript scripts from its folder; calls node; needs OPENAI_API_KEY and GEMINI_API_KEY
  • Tasks that involve Image generation

What it does

Image Gen is an agent skill from nicknisi/claude-plugins. Generate or edit images via Google Gemini (nano-banana-pro) or OpenAI gpt-image-2. Trigger on "generate image", "create diagram", "edit image", or "make illustration". Supports 1K/2K/4K resolution, masked inpainting, and text-accurate generation.

Its SKILL.md is about 500 tokens, which your agent loads only when the skill is triggered. The skill folder holds 5 other files, including scripts (for example `dist/generate_image.js`, `package.json` and `scripts/generate_image.ts`).

It sits in Media & Creative, covering Image generation and Image editing. It works with Google Gemini and OpenAI. The repository describes itself as: Nick's own marketplace of Claude Plugins. The licence is MIT.

When your agent uses it

  • Make illustration
  • Tasks that involve Image generation
  • Tasks that involve Image editing

Example prompts

  • “generate image”
  • “create diagram”
  • “edit image”
  • “/image-gen”

Requirements

  • Node.js
  • A credential in GEMINI_API_KEY
  • A credential in GOOGLE_API_KEY

What it can do on your machine

Read from SKILL.md and the folder at commit 6a6decd. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Ships 1 file in scripts/ (JavaScript and TypeScript), which the agent can run.

    Shell commands in SKILL.md call:

    • node

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names these keys or tokens, usually read from environment variables:

    • OPENAI_API_KEY
    • GEMINI_API_KEY
    • GOOGLE_API_KEY

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Image Gen loads about 502 tokens when it runs. Until then it costs about 64 tokens; SKILL.md has 166 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~64
When it runs · the whole SKILL.md, loaded when a task matches
~502

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); the scripts in this folder are not scanned.

SKILL.md

The full file from nicknisi/claude-plugins at commit 6a6decd, republished under its MIT licence (© nicknisi). 166 words, ~502 tokens.

Download SKILL.mdSave it as .claude/skills/image-gen/SKILL.md (or your agent's skills folder). This skill also uses 3 other files; get the full folder from GitHub.
name
image-gen
description
Generate or edit images via Google Gemini (nano-banana-pro) or OpenAI gpt-image-2. Trigger on "generate image", "create diagram", "edit image", or "make illustration". Supports 1K/2K/4K resolution, masked inpainting, and text-accurate generation.

Image Generation & Editing

Multi-provider image generation. Default provider is Gemini (nano-banana-pro); pass --provider openai to use gpt-image-2.

Which provider to pick

  • Gemini (default): general illustrations, quick diagrams, visual imagery. Accepts --input-image repeatedly for multi-image composition.
  • OpenAI: anything where text rendering matters (infographics, slide-like images, dense-label diagrams, logos with text), or when you need masked inpainting on edits.

Usage

The skill ships a pre-bundled Node script — no tsx or dependency install needed. The script's --help is the authoritative flag reference; run it rather than guessing:

bash
node ${CLAUDE_PLUGIN_ROOT}/skills/image-gen/dist/generate_image.js --help

Generate:

bash
node ${CLAUDE_PLUGIN_ROOT}/skills/image-gen/dist/generate_image.js \
  --prompt "your description" --filename "output.png"

Edit (image-to-image):

bash
node ${CLAUDE_PLUGIN_ROOT}/skills/image-gen/dist/generate_image.js \
  --prompt "editing instructions" --filename "output.png" \
  --input-image "path/to/input.png"

API keys

Read from the environment unless --api-key overrides: Gemini uses GEMINI_API_KEY or GOOGLE_API_KEY; OpenAI uses OPENAI_API_KEY. If only OPENAI_API_KEY is set, the provider defaults to OpenAI.

Conventions

  • Name files YYYY-MM-DD-HH-MM-SS-descriptive-name.png.
  • For blog diagrams, use OpenAI at 2K — anything with readable labels or callouts needs its text fidelity. Save to the post's content dir (e.g. src/content/blog/post-name/) and prefer clean, minimalist styles.
  • The script prints the saved path. Report that path to the user; do not read the image back.

© nicknisi, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

SKILL.md and 3 other files (scripts) in plugins/image-gen/skills/image-gen of nicknisi/claude-plugins.

  • SKILL.md
  • dist/generate_image.js
  • package.json
  • scripts/generate_image.ts

Open the folder on GitHubat commit 6a6decd

Compare with similar skills

Image Gen next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Image Gen compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Image Gen this skillnicknisi/claude-plugins114—~502Automated safety check: PassMIT
AI Image Creatorevolution-foundation/evo-nexus545—~5.1kAutomated safety check: NotesCustom licence
Generate ImageK-Dense-AI/claude-scientific-writer2.4k1 repos~3.8kAutomated safety check: NotesMIT
AI Image Generation and Editingzhayujie/CowAgent47k—~1.3kAutomated safety check: PassMIT
BlockRun Image GenerationBlockRunAI/ClawRouter6.6k—~2.1kAutomated safety check: PassMIT
Nano Banana Pro Prompts Recommend SkillYouMind-OpenLab/nano-banana-pro-prompts-recommend-skill1.9k1 repos~4.1kAutomated safety check: PassNone

Similar skills

  • AI Image Creator

    evolution-foundation/evo-nexus

    Generates PNG images through OpenRouter models, with transparent backgrounds and reference-image edits, and describes existing images with multimodal vision.

    545 GitHub stars~5.1k tokensUpdated 5 mo ago
    Media & CreativeAuto-check: notes
  • Generate Image

    K-Dense-AI/claude-scientific-writer

    Generate or edit images with AI models through the OpenRouter Image API (Gemini, Seedream, Recraft, GPT-Image, Riverflow).

    2.4k GitHub starsUsed in 1 repo~3.8k tokens
    Media & CreativeAuto-check: notes
  • Generates or edits images from text prompts through a Python script that picks an image backend based on which API keys are configured.

    47k GitHub stars~1.3k tokensUpdated today
    Media & CreativeAuto-check passed
  • BlockRun Image Generation

    BlockRunAI/ClawRouter

    Generates or edits images through ClawRouter's local image API, with a choice of models and sizes and payment handled automatically through x402.

    6.6k GitHub stars~2.1k tokensUpdated 5 days ago
    Media & CreativeAuto-check passed
  • Nano Banana Pro Prompts Recommend Skill

    YouMind-OpenLab/nano-banana-pro-prompts-recommend-skill

    Recommend suitable prompts from 10,000+ Nano Banana Pro image generation prompts based on user needs.

    1.9k GitHub starsUsed in 1 repo~4.1k tokens
    Media & CreativeAuto-check passed
  • AI Image Creator

    centminmod/my-claude-code-setup

    Generate, edit-from-reference, or analyze images with AI via OpenRouter (Gemini, GPT Image, Seedream, Qwen, MAI, Grok, FLUX.2, Recraft, Muse, Riverflow; Cloudflare AI Gateway BYOK).

    2.7k GitHub stars~8.1k tokensUpdated 2 days ago
    Media & CreativeAuto-check: notes

More from nicknisi/claude-plugins

All 12 skills in this repo
  • Blog Post Writer

    nicknisi/claude-plugins

    Turn Nick Nisi's raw material into a blog post draft in his voice.

    114 GitHub stars~1.8k tokensUpdated 2 mo ago
    Auto-check passed
  • Prototype

    nicknisi/claude-plugins

    Build a throwaway prototype to answer a design question before committing to real implementation.

    114 GitHub stars~1.2k tokensUpdated 2 mo ago
    Auto-check passed
  • Squad Review

    nicknisi/claude-plugins

    Review the current branch with six specialist lenses (security, correctness, conventions, tests, architecture, duplication), then put every finding through an adversarial verifier that tries to…

    114 GitHub stars~751 tokensUpdated 2 mo ago
    Auto-check passed
  • Tmux

    nicknisi/claude-plugins

    Read and drive other tmux panes when Claude runs inside tmux.

    114 GitHub stars~3.1k tokensUpdated 2 mo ago
    Auto-check passed
  • Youtube Notes

    nicknisi/claude-plugins

    Pull captions from a YouTube video and turn them into chapter-aligned notes where every claim carries a clickable timestamp.

    114 GitHub stars~2.8k tokensUpdated 2 mo ago
    Auto-check passed
  • Conference Talk Builder

    nicknisi/claude-plugins

    Create conference talk outlines and slide-by-slide content plans using narrative frameworks.

    114 GitHub stars~2.7k tokensUpdated 2 mo ago
    Auto-check passed

Questions about Image Gen

What does Image Gen do?

Generate or edit images via Google Gemini (nano-banana-pro) or OpenAI gpt-image-2. Image Gen is an agent skill from nicknisi/claude-plugins. Generate or edit images via Google Gemini (nano-banana-pro) or OpenAI gpt-image-2.

When should I use Image Gen?

Image Gen fits situations like: make illustration; tasks that involve Image generation; tasks that involve Image editing.

How do I install Image Gen in Claude Code?

Run `npx skills add nicknisi/claude-plugins --skill image-gen -a claude-code`. Or copy the skill folder (plugins/image-gen/skills/image-gen in nicknisi/claude-plugins) into .claude/skills/image-gen in your project. Claude Code loads it when a task matches its description.

How do I install Image Gen in Codex?

Run `npx skills add nicknisi/claude-plugins --skill image-gen -a codex`. Or copy the skill folder (plugins/image-gen/skills/image-gen in nicknisi/claude-plugins) into .agents/skills/image-gen in your project. Codex loads it when a task matches its description.

Can I use Image Gen in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add nicknisi/claude-plugins --skill image-gen -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/image-gen, .gemini/skills/image-gen, .github/skills/image-gen and .opencode/skills/image-gen in your project.

What does Image Gen need to run?

Going by SKILL.md and its folder, Image Gen needs JavaScript and TypeScript for the scripts in its folder, the command-line tools its instructions call (node) and credentials named OPENAI_API_KEY, GEMINI_API_KEY and GOOGLE_API_KEY. Our summary lists: Node.js; A credential in GEMINI_API_KEY; A credential in GOOGLE_API_KEY.

Does Image Gen access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is Image Gen safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. The check reads SKILL.md only: the scripts in the folder are not scanned, so read them before running anything.

What licence does Image Gen use?

Image Gen is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Image Gen use?

About 502 tokens (SKILL.md is roughly 2k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Image Gen?

Skills that share tags, products or a category with Image Gen: AI Image Creator (evolution-foundation/evo-nexus, 545 stars), Generate Image (K-Dense-AI/claude-scientific-writer, 2.4k stars), AI Image Generation and Editing (zhayujie/CowAgent, 47k stars) and BlockRun Image Generation (BlockRunAI/ClawRouter, 6.6k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Image Gen?

nicknisi (a GitHub user) maintains it in nicknisi/claude-plugins, which has 114 GitHub stars. The repository holds 12 skills in this directory. The repository was last updated on August 11, 2026.

Source: nicknisi/claude-plugins on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.