Agent skill

Invokeai Image Gen

by sammcj in sammcj/agentic-coding

Generate images using InvokeAI's local API. An agent skill from sammcj/agentic-coding.

Apache-2.0Auto-check passedMedia & Creative

Install Invokeai Image Gen

skills CLI
$ npx skills add sammcj/agentic-coding --skill invokeai-image-gen -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install sammcj/agentic-coding invokeai-image-gen --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/sammcj/agentic-coding.git skills-src && mkdir -p .claude/skills && cp -r skills-src/Skills_disabled/invokeai-image-gen .claude/skills/invokeai-image-gen && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
invokeai-image-gen
GitHub stars
162
Token cost
~1.2k tokens
SKILL.md length
436 words
Files
2 (incl. scripts)
Skills in repo
64
Repo updated
First seen
Licence
Apache-2.0

At a glance

Generate images using InvokeAI's local API. An agent skill from sammcj/agentic-coding.

  • Works in 3 steps: Front-load critical elements (word order… → Specify lighting: source, quality,… → Include sensory texture: materials,…
  • Asked to generate
  • SKILL.md covers Quick Start, Overriding The Default Model, Options and Model Defaults, plus 4 more sections
  • Runs Python scripts from its folder; calls python; needs MODEL_KEY and INVOKEAI_AUTH_TOKEN

What it does

Invokeai Image Gen is an agent skill from sammcj/agentic-coding. Generate images using InvokeAI's local API. Use when asked to generate, create, or make images with InvokeAI, FLUX.2 Klein, Z-Image Turbo, FLUX, or SDXL models. Supports text-to-image generation, automatic model detection, image download, and parameter selection based on model architecture.

Its SKILL.md is about 1.2k tokens, which your agent loads only when the skill is triggered. The skill folder holds 2 other files, including scripts (for example `scripts/generate.py`).

It sits in Media & Creative, covering Image generation. It works with Stable Diffusion. The repository describes itself as: Agentic Coding Rules, Templates etc... The licence is Apache-2.0.

When your agent uses it

  • Asked to generate
  • Make images with InvokeAI

Example prompts

  • “/invokeai-image-gen”

Requirements

  • Python 3
  • A credential in MODEL_KEY
  • A credential in INVOKEAI_AUTH_TOKEN

Workflow steps

3 steps, taken from the first numbered list in SKILL.md.

  1. Front-load critical elements (word order matters)
  2. Specify lighting: source, quality, direction, temperature
  3. Include sensory texture: materials, reflections, atmosphere

What it can do on your machine

Read from SKILL.md and the folder at commit 2f25ced. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Ships 1 file in scripts/ (Python), which the agent can run.

    Shell commands in SKILL.md call:

    • python

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names these keys or tokens, usually read from environment variables:

    • MODEL_KEY
    • INVOKEAI_AUTH_TOKEN

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Invokeai Image Gen loads about 1.2k tokens when it runs. Until then it costs about 78 tokens; SKILL.md has 436 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~78
When it runs · the whole SKILL.md, loaded when a task matches
~1.2k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); the scripts in this folder are not scanned.

SKILL.md

The full file from sammcj/agentic-coding at commit 2f25ced, republished under its Apache-2.0 licence (© sammcj). 436 words, ~1,183 tokens.

Download SKILL.mdSave it as .claude/skills/invokeai-image-gen/SKILL.md (or your agent's skills folder). This skill also uses 1 other file; get the full folder from GitHub.
name
invokeai-image-gen
description
Generate images using InvokeAI's local API. Use when asked to generate, create, or make images with InvokeAI, FLUX.2 Klein, Z-Image Turbo, FLUX, or SDXL models. Supports text-to-image generation, automatic model detection, image download, and parameter selection based on model architecture.

InvokeAI Image Generation

Generate images via InvokeAI's REST API. Supports FLUX.2 Klein (default), Z-Image Turbo, FLUX.1, and SDXL.

Quick Start

Simply call the script with your prompt and the output file name:

bash
python scripts/generate.py -p "A dramatic sunset over snow-capped mountains, warm orange light reflecting off a still alpine lake in the foreground. Soft clouds catch the fading light." -o sunset.png

Overriding The Default Model

If the user asks you to use a specific model, first find the model key, then use it in the command:

bash
python scripts/generate.py --list-models | grep -i 'flux'

python scripts/generate.py -p "A tabby cat with bright green eyes sits on a weathered wooden windowsill, soft afternoon light streaming through lace curtains. Cosy, intimate mood." --model MODEL_KEY -o cat.png

Options

OptionDescription
--prompt, -pGeneration prompt (required)
--negative, -nNegative prompt (SDXL only)
--model, -mModel key (UUID) or partial name match
--width, -W / --height, -HDimensions
--steps, -sDenoising steps
--cfg, -cCFG scale
--guidance, -gGuidance strength (FLUX.1 only)
--schedulerSampling scheduler
--seedRandom seed
--output, -oOutput path (default: invokeai-{seed}.png)
--list-modelsList installed models
--jsonJSON output

Model Defaults

Note: FLUX.2 Klein is the latest model which is used by default.

ModelStepsGuidanceCFGScheduler
FLUX.2 Klein43.51.0euler
Z-Image Turbo9-1.0euler
FLUX.1 dev283.51.0euler
FLUX.1 Krea dev284.51.0euler
FLUX.1 Kontext dev282.51.0euler
FLUX.1 schnell40.01.0euler
SDXL25-6.0dpmpp_2m_k
SDXL Turbo8-1.0dpmpp_sde

All models default to 1024x1024. FLUX requires dimensions divisible by 16, SDXL by 8.

FLUX.1 Variant Notes
  • FLUX.1 dev: Standard text-to-image model, balanced quality/speed
  • FLUX.1 Krea dev: Fine-tuned for aesthetic photography, use higher guidance (4.5)
  • FLUX.1 Kontext dev: Image editing model, use lower guidance (2.5)
  • FLUX.1 schnell: Distilled fast model, 4 steps, no guidance needed
Show full SKILL.md (198 more words)Show less

Model Selection

Auto-priority: Klein > Z-Image > FLUX > SDXL

Detection by name/base:

  • flux2_klein: "klein" in name or "flux2" in base
  • flux_krea: "krea" in name (FLUX.1 base)
  • flux_kontext: "kontext" in name (FLUX.1 base)
  • flux_schnell: "schnell" in name (FLUX.1 base)
  • flux: "flux" in base (standard dev)
  • zimage: "z-image" in base or "z-image/zimage" in name
  • sdxl: "sdxl" in base (turbo/lightning variants auto-detect)

Prompting (general information, but especially useful for FLUX.2 Klein)

Write prose, not keywords. Structure: Subject -> Setting -> Details -> Lighting -> Atmosphere

A weathered fisherman in his late sixties stands at the bow of a wooden boat,
wearing a salt-stained wool sweater. Golden hour sunlight filters through
morning mist, creating quiet determination and solitude.

Key techniques:

  1. Front-load critical elements (word order matters)
  2. Specify lighting: source, quality, direction, temperature
  3. Include sensory texture: materials, reflections, atmosphere

Good: "A woman with short blonde hair poses against a light neutral background wearing colourful earrings, resting her chin on her hand."

Bad: "woman, blonde, short hair, neutral background, earrings"

Append style tags: Style: Country chic. Mood: Serene, romantic.

Troubleshooting

IssueSolution
Connection refusedCheck InvokeAI is running
Model not foundUse --list-models for valid keys
Dimensions errorFLUX: multiples of 16, SDXL: 8
Black images (macOS)Set precision: bfloat16 in invokeai.yaml

If the script fails to find the URL or authentication token, you can set or ask the user to set environment variables:

bash
export INVOKEAI_API_URL='http://localhost:9090'
export INVOKEAI_AUTH_TOKEN='your-token'  # Optional

Resources

  • scripts/generate.py - Main generation script

© sammcj, Apache-2.0. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

SKILL.md and 1 other file (scripts) in Skills_disabled/invokeai-image-gen of sammcj/agentic-coding.

  • SKILL.md
  • scripts/generate.py

Open the folder on GitHubat commit 2f25ced

Compare with similar skills

Invokeai Image Gen next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Invokeai Image Gen compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Invokeai Image Gen this skillsammcj/agentic-coding162—~1.2kAutomated safety check: PassApache-2.0
Stable Diffusion with DiffusersOrchestra-Research/AI-Research-SKILLs13k5 repos~3.2kAutomated safety check: PassMIT
Stability AIsickn33/agentic-awesome-skills47k2 repos~2kAutomated safety check: NotesMIT
Image Genamd/gaia1.6k—~1.4kAutomated safety check: PassMIT
Stability AIdiegosouzapw/awesome-omni-skills159—~4.1kAutomated safety check: NotesMIT
Nano Banana Pro Prompts Recommend SkillYouMind-OpenLab/nano-banana-pro-prompts-recommend-skill1.9k1 repos~4.1kAutomated safety check: PassNone

Similar skills

  • Stable Diffusion with Diffusers

    Orchestra-Research/AI-Research-SKILLs

    Generates and edits images with Stable Diffusion through Hugging Face Diffusers, covering text-to-image, image-to-image, inpainting, SDXL and custom pipelines.

    13k GitHub starsUsed in 5 repos~3.2k tokens
    Media & CreativeAuto-check passed
  • Stability AI

    sickn33/agentic-awesome-skills

    Geracao de imagens via Stability AI (SD3.5, Ultra, Core). An agent skill from sickn33/agentic-awesome-skills.

    47k GitHub starsUsed in 2 repos~2k tokens
    Media & CreativeAuto-check: notes
  • Image Gen

    amd/gaia

    Turn a description into an image file with local Stable Diffusion, then iterate on it.

    1.6k GitHub stars~1.4k tokensUpdated today
    Media & CreativeAuto-check passed
  • Stability AI

    diegosouzapw/awesome-omni-skills

    Stability AI — Gerador de Imagens Profissional workflow skill.

    159 GitHub stars~4.1k tokensUpdated 3 mo ago
    Media & CreativeAuto-check: notes
  • Nano Banana Pro Prompts Recommend Skill

    YouMind-OpenLab/nano-banana-pro-prompts-recommend-skill

    Recommend suitable prompts from 10,000+ Nano Banana Pro image generation prompts based on user needs.

    1.9k GitHub starsUsed in 1 repo~4.1k tokens
    Media & CreativeAuto-check passed
  • Iib

    zanllp/infinite-image-browsing

    Interact with IIB (Infinite Image Browsing) service for searching, browsing, tagging, and organizing AI-generated images.

    1.4k GitHub stars~3.3k tokensUpdated today
    Media & CreativeAuto-check passed

More from sammcj/agentic-coding

All 64 skills in this repo
  • Yue2 Music

    sammcj/agentic-coding

    A skill your agent uses when generating songs with YuE2, covering a recording via SheetSage2 audio-to-ABC, editing a score or lyrics with melody preservation, or building a reproducible listening…

    162 GitHub stars~2.3k tokensUpdated yesterday
    Auto-check passed
  • Bento Slides

    sammcj/agentic-coding

    A skill your agent uses when creating or editing Bento (.bento.html) slide decks, including any request for a single-file HTML slide deck.

    162 GitHub stars~2.9k tokensUpdated yesterday
    Auto-check passed
  • Idrive Backup

    sammcj/agentic-coding

    A skill your agent uses whenever the user wants you to manage, discuss or diagnose iDrive Backup configuration on macOS

    162 GitHub stars~1.7k tokensUpdated yesterday
    Auto-check: notes
  • Piper Tts Training

    sammcj/agentic-coding

    Train custom TTS voices for Piper (ONNX format) using fine-tuning or from-scratch approaches.

    162 GitHub stars~1.4k tokensUpdated yesterday
    Auto-check passed
  • PPTX To Md

    sammcj/agentic-coding

    Convert a PPTX slide deck into per-slide markdown that preserves both the verbatim text and the meaning of embedded screenshots, diagrams and charts in their original layout positions.

    162 GitHub stars~1.8k tokensUpdated yesterday
    Auto-check passed
  • Skill Creator Primer

    sammcj/agentic-coding

    You MUST load this skill before the skill-creator skill AND before making ANY change to, or conducting a review of ANY Agent Skill.

    162 GitHub stars~9.8k tokensUpdated yesterday
    Auto-check passed

Questions about Invokeai Image Gen

What does Invokeai Image Gen do?

Generate images using InvokeAI's local API. An agent skill from sammcj/agentic-coding. Invokeai Image Gen is an agent skill from sammcj/agentic-coding. Generate images using InvokeAI's local API.

When should I use Invokeai Image Gen?

Invokeai Image Gen fits situations like: asked to generate; make images with InvokeAI.

How do I install Invokeai Image Gen in Claude Code?

Run `npx skills add sammcj/agentic-coding --skill invokeai-image-gen -a claude-code`. Or copy the skill folder (Skills_disabled/invokeai-image-gen in sammcj/agentic-coding) into .claude/skills/invokeai-image-gen in your project. Claude Code loads it when a task matches its description.

How do I install Invokeai Image Gen in Codex?

Run `npx skills add sammcj/agentic-coding --skill invokeai-image-gen -a codex`. Or copy the skill folder (Skills_disabled/invokeai-image-gen in sammcj/agentic-coding) into .agents/skills/invokeai-image-gen in your project. Codex loads it when a task matches its description.

Can I use Invokeai Image Gen in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add sammcj/agentic-coding --skill invokeai-image-gen -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/invokeai-image-gen, .gemini/skills/invokeai-image-gen, .github/skills/invokeai-image-gen and .opencode/skills/invokeai-image-gen in your project.

What does Invokeai Image Gen need to run?

Going by SKILL.md and its folder, Invokeai Image Gen needs Python for the scripts in its folder, the command-line tools its instructions call (python) and credentials named MODEL_KEY and INVOKEAI_AUTH_TOKEN. Our summary lists: Python 3; A credential in MODEL_KEY; A credential in INVOKEAI_AUTH_TOKEN.

Does Invokeai Image Gen access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is Invokeai Image Gen safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. The check reads SKILL.md only: the scripts in the folder are not scanned, so read them before running anything.

What licence does Invokeai Image Gen use?

Invokeai Image Gen is published under the Apache-2.0 licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Invokeai Image Gen use?

About 1.2k tokens (SKILL.md is roughly 4.7k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Invokeai Image Gen?

Skills that share tags, products or a category with Invokeai Image Gen: Stable Diffusion with Diffusers (Orchestra-Research/AI-Research-SKILLs, 13k stars), Stability AI (sickn33/agentic-awesome-skills, 47k stars), Image Gen (amd/gaia, 1.6k stars) and Stability AI (diegosouzapw/awesome-omni-skills, 159 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Invokeai Image Gen?

sammcj (a GitHub user) maintains it in sammcj/agentic-coding, which has 162 GitHub stars. The repository holds 64 skills in this directory. The repository was last updated on October 9, 2026.

Source: sammcj/agentic-coding on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.