Agent skill

Imagen

by sanjay3290 in sanjay3290/ai-skills

Generate images using Google Gemini's image generation capabilities.

Apache-2.0Auto-check passedMedia & Creative

Install Imagen

skills CLI
$ npx skills add sanjay3290/ai-skills --skill imagen -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install sanjay3290/ai-skills imagen --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/sanjay3290/ai-skills.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/imagen .claude/skills/imagen && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
imagen
GitHub stars
431
Used in
6 other repos
Token cost
~657 tokens
SKILL.md length
194 words
Files
7 (incl. scripts)
Skills in repo
24
Repo updated
First seen
Licence
Apache-2.0

At a glance

Generate images using Google Gemini's image generation capabilities.

  • Works in 4 steps: Takes a text prompt describing the… → Calls Google Gemini API with image… → Saves the generated image to a specified… → …
  • The user needs to create
  • SKILL.md covers Overview, When to Use This Skill, How It Works and Usage, plus 3 more sections
  • Runs Python scripts from its folder; calls python; needs GEMINI_API_KEY

What it does

Imagen is an agent skill from sanjay3290/ai-skills. Generate images using Google Gemini's image generation capabilities. Use this skill when the user needs to create, generate, or produce images for any purpose including UI mockups, icons, illustrations, diagrams, concept art, placeholder images, or visual representations.

Its SKILL.md is about 660 tokens, which your agent loads only when the skill is triggered. The skill folder holds 7 other files, including scripts (for example `README.md`, `examples.md` and `reference.md`).

It sits in Media & Creative, covering Image generation, Diagrams and Icons and illustration. It works with Google Gemini. The repository describes itself as: 24 cross-platform agent skills for Claude Code, Cursor, Codex & Gemini CLI — databases, messaging, research, TTS, DevOps, and Google Workspace. The licence is Apache-2.0.

When your agent uses it

  • The user needs to create
  • Produce images for any purpose including UI mockups
  • Placeholder images
  • Visual representations

Example prompts

  • “/imagen”

Requirements

  • Python 3
  • A credential in GEMINI_API_KEY

Workflow steps

4 steps, taken from the first numbered list in SKILL.md.

  1. Takes a text prompt describing the desired image
  2. Calls Google Gemini API with image generation configuration
  3. Saves the generated image to a specified location (defaults to current directory)
  4. Returns the file path for use in your project

What it can do on your machine

Read from SKILL.md and the folder at commit 281d88d. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Ships 1 file in scripts/ (Python), which the agent can run.

    Shell commands in SKILL.md call:

    • python

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names these keys or tokens, usually read from environment variables:

    • GEMINI_API_KEY

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Imagen loads about 657 tokens when it runs. Until then it costs about 70 tokens; SKILL.md has 194 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~70
When it runs · the whole SKILL.md, loaded when a task matches
~657

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); the scripts in this folder are not scanned.

SKILL.md

The full file from sanjay3290/ai-skills at commit 281d88d, republished under its Apache-2.0 licence (© sanjay3290). 194 words, ~657 tokens.

Download SKILL.mdSave it as .claude/skills/imagen/SKILL.md (or your agent's skills folder). This skill also uses 6 other files; get the full folder from GitHub.
name
imagen
description
Generate images using Google Gemini's image generation capabilities. Use this skill when the user needs to create, generate, or produce images for any purpose including UI mockups, icons, illustrations, diagrams, concept art, placeholder images, or visual representations.
license
Apache-2.0
metadata.author
sanjay3290
metadata.version
1.0

Imagen - AI Image Generation Skill

Overview

This skill generates images using Google Gemini's image generation model (gemini-3-pro-image-preview). It enables seamless image creation during any Claude Code session - whether you're building frontend UIs, creating documentation, or need visual representations of concepts.

Cross-Platform: Works on Windows, macOS, and Linux.

When to Use This Skill

Automatically activate this skill when:

  • User requests image generation (e.g., "generate an image of...", "create a picture...")
  • Frontend development requires placeholder or actual images
  • Documentation needs illustrations or diagrams
  • Visualizing concepts, architectures, or ideas
  • Creating icons, logos, or UI assets
  • Any task where an AI-generated image would be helpful

How It Works

  1. Takes a text prompt describing the desired image
  2. Calls Google Gemini API with image generation configuration
  3. Saves the generated image to a specified location (defaults to current directory)
  4. Returns the file path for use in your project

Usage

bash
# Basic usage
python scripts/generate_image.py "A futuristic city skyline at sunset"

# With custom output path
python scripts/generate_image.py "A minimalist app icon for a music player" "./assets/icons/music-icon.png"

# With custom size
python scripts/generate_image.py --size 2K "High resolution landscape" "./wallpaper.png"

Requirements

  • GEMINI_API_KEY environment variable must be set
  • Python 3.6+ (uses standard library only, no pip install needed)

Output

Generated images are saved as PNG files. The script returns:

  • Success: Path to the generated image
  • Failure: Error message with details

Examples

Frontend Development
User: "I need a hero image for my landing page - something abstract and tech-focused"
-> Generates and saves image, provides path for use in HTML/CSS
Documentation
User: "Create a diagram showing microservices architecture"
-> Generates visual representation, ready for README or docs
UI Assets
User: "Generate a placeholder avatar image for the user profile component"
-> Creates image in appropriate size for component use

© sanjay3290, Apache-2.0. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

SKILL.md and 6 other files (scripts) in skills/imagen of sanjay3290/ai-skills.

  • SKILL.md
  • .env.example
  • .gitignore
  • README.md
  • examples.md
  • reference.md
  • scripts/generate_image.py

Open the folder on GitHubat commit 281d88d

Used in 6 other repositories

We found 15 copies of this SKILL.md (exact, near-identical or edited) in other folders, from 6 other GitHub owners. This page covers the copy in sanjay3290/ai-skills, which our catalogue first saw on October 7, 2026.

Compare with similar skills

Imagen next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Imagen compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Imagen this skillsanjay3290/ai-skills4316 repos~657Automated safety check: PassApache-2.0
Cursor Image Generationtmcfarlane/oh-my-cursor110—~1.8kAutomated safety check: PassMIT
Sf Diagram NanobananaproJaganpro/sf-skills424—~1.6kAutomated safety check: PassMIT
Imagennexu-io/open-design100k—~279Automated safety check: PassApache-2.0
Image PromptingBlockRunAI/blockrun-mcp391—~3.6kAutomated safety check: PassMIT
Imagegennexu-io/open-design100k—~300Automated safety check: PassApache-2.0

Similar skills

  • Cursor Image Generation

    tmcfarlane/oh-my-cursor

    Generate and iterate images in Cursor using the built-in image model and strong prompts.

    110 GitHub stars~1.8k tokensUpdated 3 mo ago
    Media & CreativeAuto-check passed
  • Sf Diagram Nanobananapro

    Jaganpro/sf-skills

    AI-powered image generation for Salesforce visuals via Nano Banana Pro.

    424 GitHub stars~1.6k tokensUpdated 5 mo ago
    Media & CreativeAuto-check passed
  • Imagen

    nexu-io/open-design

    Generate images using Google Gemini's image generation API for UI mockups, icons, illustrations, and visual assets.

    100k GitHub stars~279 tokensUpdated today
    Media & CreativeAuto-check passed
  • Image Prompting

    BlockRunAI/blockrun-mcp

    A skill your agent uses when generating or editing images via blockrunimage — especially with GPT Image 2, Nano Banana, or Grok Imagine for posters, UI mockups, marketing assets, product shots, or…

    391 GitHub stars~3.6k tokensUpdated yesterday
    Media & CreativeAuto-check passed
  • Imagegen

    nexu-io/open-design

    Generate and edit images using OpenAI's Image API for project assets — UI mockups, icons, illustrations, social cards, and visual references.

    100k GitHub stars~300 tokensUpdated today
    Media & CreativeAuto-check passed
  • Nano Banana

    kkoppenhaver/cc-nano-banana

    REQUIRED for all image generation requests. An agent skill from kkoppenhaver/cc-nano-banana.

    378 GitHub starsUsed in 1 repo~1.4k tokens
    Media & CreativeAuto-check passed

More from sanjay3290/ai-skills

All 24 skills in this repo
  • Deep Research

    sanjay3290/ai-skills

    Execute autonomous multi-step research using Google Gemini Deep Research Agent.

    431 GitHub starsUsed in 9 repos~683 tokens
    Auto-check: notes
  • Notebooklm

    sanjay3290/ai-skills

    Query and manage Google NotebookLM notebooks with persistent profile auth, source sync, batch/multi queries, and structured exports.

    431 GitHub stars~655 tokensUpdated 29 days ago
    Auto-check passed
  • Whatsapp

    sanjay3290/ai-skills

    Send and receive WhatsApp messages via the unofficial linked-device client pywhats (pip install pywhats) — pair with QR, send text/images, group chat, read receipts, presence/typing, and a…

    431 GitHub stars~1.1k tokensUpdated 29 days ago
    Auto-check passed
  • Postgres

    sanjay3290/ai-skills

    Execute read-only SQL queries against multiple PostgreSQL databases.

    431 GitHub starsUsed in 1 repo~975 tokens
    Auto-check passed
  • Atlassian

    sanjay3290/ai-skills

    Manage Jira issues and Confluence wiki pages in Atlassian Cloud.

    431 GitHub stars~1.6k tokensUpdated 29 days ago
    Auto-check passed
  • Azure Devops

    sanjay3290/ai-skills

    Manage Azure DevOps projects, work items, repos, PRs, pipelines, wikis, test plans, security alerts, variable groups, environments/approvals, branch policies, and attachments.

    431 GitHub stars~4.2k tokensUpdated 29 days ago
    Auto-check passed

Works with

Questions about Imagen

What does Imagen do?

Generate images using Google Gemini's image generation capabilities. Imagen is an agent skill from sanjay3290/ai-skills. Generate images using Google Gemini's image generation capabilities.

When should I use Imagen?

Imagen fits situations like: the user needs to create; produce images for any purpose including UI mockups; placeholder images; visual representations.

How do I install Imagen in Claude Code?

Run `npx skills add sanjay3290/ai-skills --skill imagen -a claude-code`. Or copy the skill folder (skills/imagen in sanjay3290/ai-skills) into .claude/skills/imagen in your project. Claude Code loads it when a task matches its description.

How do I install Imagen in Codex?

Run `npx skills add sanjay3290/ai-skills --skill imagen -a codex`. Or copy the skill folder (skills/imagen in sanjay3290/ai-skills) into .agents/skills/imagen in your project. Codex loads it when a task matches its description.

Can I use Imagen in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add sanjay3290/ai-skills --skill imagen -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/imagen, .gemini/skills/imagen, .github/skills/imagen and .opencode/skills/imagen in your project.

What does Imagen need to run?

Going by SKILL.md and its folder, Imagen needs Python for the scripts in its folder, the command-line tools its instructions call (python) and credentials named GEMINI_API_KEY. Our summary lists: Python 3; A credential in GEMINI_API_KEY.

Does Imagen access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is Imagen safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. The check reads SKILL.md only: the scripts in the folder are not scanned, so read them before running anything.

What licence does Imagen use?

Imagen is published under the Apache-2.0 licence (declared in SKILL.md). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Imagen use?

About 657 tokens (SKILL.md is roughly 2.6k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Imagen?

Skills that share tags, products or a category with Imagen: Cursor Image Generation (tmcfarlane/oh-my-cursor, 110 stars), Sf Diagram Nanobananapro (Jaganpro/sf-skills, 424 stars), Imagen (nexu-io/open-design, 100k stars) and Image Prompting (BlockRunAI/blockrun-mcp, 391 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Imagen?

sanjay3290 (a GitHub user) maintains it in sanjay3290/ai-skills, which has 431 GitHub stars. The repository holds 24 skills in this directory. The repository was last updated on September 10, 2026.

Source: sanjay3290/ai-skills on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.