Agent skill

Image

by smixs in smixs/visual-skills

Image prompting skill for Nano Banana (NBP/NB2) and GPT Image 2.5 (Flare/Sunburst).

CC-BY-4.0Auto-check passedMedia & Creative

Install Image

skills CLI
$ npx skills add smixs/visual-skills --skill image -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install smixs/visual-skills image --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/smixs/visual-skills.git skills-src && mkdir -p .claude/skills && cp -r skills-src/image .claude/skills/image && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
image
GitHub stars
484
Token cost
~2.1k tokens
SKILL.md length
851 words
Files
24 (incl. references)
Skills in repo
2
Repo updated
First seen
Licence
CC-BY-4.0

At a glance

Image prompting skill for Nano Banana (NBP/NB2) and GPT Image 2.5 (Flare/Sunburst).

  • Works in 6 steps: always read first → models.md → read one model file (the one you picked) → always read after the model file →… → …
  • Сгенерируй картинку
  • Instructions only: no scripts, shell commands, URLs or credentials in SKILL.md
  • Промпт для картинки

What it does

Image is an agent skill from smixs/visual-skills. Image prompting skill for Nano Banana (NBP/NB2) and GPT Image 2.5 (Flare/Sunburst). Writes ready-to-use prompts with model/quality/size recommendations. Use when: "нарисуй", "сгенерируй картинку", "image prompt", "промпт для картинки", blog covers, slides, posters, product shots, UI mockups, storyboards, character sheets, edit/colorize, style transfer, vision analysis, image-to-prompt, nb, NBP, NB2, gpt-image-2.5, multi-panel grids, ecommerce product photography, fashion editorial, food/beverage ads, cinematic…

Its SKILL.md is about 2.1k tokens, which your agent loads only when the skill is triggered. The skill folder holds 25 other files, including reference files (for example `references/characters.md`, `references/creative-direction.md` and `references/de-slop.md`).

It sits in Media & Creative, covering Image generation and AI video generation. It works with Google Gemini, Seedance and Claude Agent SDK. The repository describes itself as: AI film director skills for agents: cinematic dramaturgy (Murch, blocking, montage) + exact prompt syntax for Seedance 2.5, Kling 3.0 Turbo/Omni, Veo 3.1, Nano Banana 2, GPT… The licence is CC-BY-4.0.

When your agent uses it

  • Сгенерируй картинку
  • Промпт для картинки
  • Character sheets
  • Vision analysis

Example prompts

  • “image prompt”
  • “/image”

Workflow steps

6 steps, taken from the step headings in SKILL.md.

  1. always read first → models.md
  2. read one model file (the one you picked)
  3. always read after the model file → golden-rules.md
  4. task-shaped reading (load only what matches the request)
  5. read for production language → creative-direction.md
  6. read if structuring a complex prompt → prompt-framework.md

What it can do on your machine

Read from SKILL.md and the folder at commit 92be33a. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    No scripts in the folder and no shell commands in SKILL.md.

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    Links to these hosts (documentation or services it may open):

    • github.com
    • t.me
    • sergeshima.com
    • aimasters.me

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Image loads about 2.1k tokens when it runs, and up to ~58k if it reads all its reference files. Until then it costs about 152 tokens; SKILL.md has 851 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~152
When it runs · the whole SKILL.md, loaded when a task matches
~2.1k
With references · SKILL.md plus every file in references/, read only if the agent opens them
~58k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from smixs/visual-skills at commit 92be33a, republished under its CC-BY-4.0 licence (© smixs). 851 words, ~2,149 tokens.

Download SKILL.mdSave it as .claude/skills/image/SKILL.md (or your agent's skills folder). This skill also uses 23 other files; get the full folder from GitHub.
name
image
description
Image prompting skill for Nano Banana (NBP/NB2) and GPT Image 2.5 (Flare/Sunburst). Writes ready-to-use prompts with model/quality/size recommendations. Use when: "нарисуй", "сгенерируй картинку", "image prompt", "промпт для картинки", blog covers, slides, posters, product shots, UI mockups, storyboards, character sheets, edit/colorize, style transfer, vision analysis, image-to-prompt, nb, NBP, NB2, gpt-image-2.5, multi-panel grids, ecommerce product photography, fashion editorial, food/beverage ads, cinematic portraits. Do NOT use for: video (use video skill), 3D models, audio, non-image tasks.
license
CC-BY-4.0 (attribution required — Serge Shima, github.com/smixs/visual-skills)

Image Prompting — Nano Banana & GPT Image 2.5

This skill writes image prompts. It does not generate images. The output is: model name + quality / size / aspect ratio + the prompt itself.

The body of this SKILL.md is intentionally thin so you cannot fake a result by reading it alone. The actual rules — what the models reward, what they punish, how to phrase a 5-slot template, when to add quality: high, when to use image grounding — live only in the reference files.

Route first — is this actually an image-prompt task?

  • Motion, clips, montage (Seedance, Kling, Veo, any image-to-video): use the sibling video skill. This skill's storyboard and keyframe outputs feed it.
  • No idea or script yet (user wants a concept or an ad scenario, not a picture): if the creative-director skill is installed, start there — it develops ideas and scripts for commercials and beyond (github.com/smixs/creative-director-skill).
  • A concrete image is needed — this skill. Continue below.

Mandatory reading order — DO NOT WRITE A PROMPT WITHOUT THIS

Past attempts to write prompts directly from this skill body produced lazy, generic results. Each model has its own physics; common rules collapse into mush when applied without model-specific syntax. Read in this order before producing any prompt:

Step 1 — always read first → models.md

Decide: Nano Banana (NB2 or NBP) or GPT Image 2.5 (Flare for speed, Sunburst for precision edits). The choice changes the prompt syntax fundamentally — natural-language paragraphs vs. labeled 5-slot template, quality settings, which features exist (image grounding only on NB, EXACT TEXT discipline only on GPT Image, etc.).

If the user named a model — confirm and proceed. If not — pick using the table in models.md, then state your choice in the output header.

Step 2 — read one model file (the one you picked)
  • Nano Banana → nano-banana.md Image grounding for real locations. Extreme aspect ratios (1:8, 8:1, 4:1). Thinking mode. JSON for 5+ elements. Up to 14 reference images. Why you must NOT write 50mm / f-stop / ISO numbers.

  • GPT Image 2.5 → gpt-image.md 5-slot template (Scene / Subject / Important Details / Use Case / Constraints). Anti-slop banned-words list. quality: low / medium / high / xhigh / max as a deliberate fidelity lever. Size constraints (multiples of 16, max 3:1, up to 4K 3840×2160). Two-column edit logic (Change / Preserve / Constraints). Up to 16 reference images with explicit roles.

The model file is non-negotiable. Skipping it is the single biggest cause of weak prompts.

Step 3 — always read after the model file → golden-rules.md

Universal rules that apply to both models: start with a verb, positive framing, hex colors, quote text, edit don't re-roll, one change per iteration, reference images.

Show full SKILL.md (427 more words)Show less
Step 4 — task-shaped reading (load only what matches the request)

Pick zero or more, depending on what the user asked for:

  • Text in image, infographic, diagram, multilingual rendering → text-rendering.md
  • Edit existing image (object removal, lighting swap, colorization, restoration, localization) → editing.md
  • Character continuity across multiple images / panels → characters.md
  • The image must pass as a real photograph (portrait, reportage, UGC, casting, product-in-hand) — or the user says the result "looks AI", "too glossy", "not like the reference" → de-slop.md. Model default priors, banned booster words, capture pipeline instead of adjectives, located imperfections.
  • Presentation slides → slides.md
  • Sequential narrative (storyboard, comic, panel sequence) → storyboards.md
  • Sketch → final, wireframes, structural input → structural.md
  • 2D → 3D, floor plans, isometric → dimensional.md
  • Vision analysis / image-to-prompt / style transfer from a reference image → vision-decomposer.md. Load this whenever the user attaches an image and asks to recreate, match, decompose, or transfer its style.
  • Multi-panel compositions (grids, collages, storyboard sheets in ONE image) → multi-panel.md. 9-cell TVC grids, 2x2 portrait grids, 3-panel campaign collages, 4x3 borderless grids, 6-frame cinematic sequences, before/after splits, 12-panel storyboard posters.
  • Industry pattern libraries — proven prompt templates by vertical. Load the matching file:
Step 5 — read for production language → creative-direction.md

Studio-quality vocabulary for lighting design, camera and hardware, color grading and film stock, materiality and texture. Read when you need precise terms beyond what golden-rules.md covers.

Step 6 — read if structuring a complex prompt → prompt-framework.md

Universal element checklist (subject, context, action, environment, camera, lighting, mood, materials, palette, format), detail modes (concise / standard / verbose / cinematic verbose), parameterized templates, output structure with parameters and exclusions.


Output format

When you return the prompt, structure it like this:

Model: <nano-banana-2 | nano-banana-pro | gpt-image-2.5-flare | gpt-image-2.5-sunburst>
Quality: <low | medium | high | xhigh | max>   (only for gpt-image-2.5)
Size / Ratio: <e.g. 1536×1024 or 16:9>

Prompt:
<the prompt text, ready to copy>

Notes:
- <anything you inferred or assumed because the user did not specify>

For edits, also include an explicit preserve-list (mandatory for gpt-image-2.5, recommended for nano-banana):

Change: <one concrete thing>
Preserve: <face, pose, lighting, framing, geometry, ...>
Constraints: <no extra objects, no drift, ...>

Final response style

Prefer: ready-to-copy prompts, hex colors, concrete materials, named compositions, model-specific syntax (5-slot for GPT Image, natural prose for Nano Banana).

Avoid: tag soup ("cool, modern, 4k"), vague praise ("stunning, epic, masterpiece" — actively hurts GPT Image 2.5), negative framing ("no people, no cars" — invert to positive), external comparisons ("like Apple ad" — describe the visual properties instead), numerical lens parameters in Nano Banana prompts (it ignores them).


Author: Serge Shima (t.me/aimastersme · sergeshima.com · aimasters.me) · License: CC BY 4.0 — attribution required · Source: smixs/visual-skills

© smixs, CC-BY-4.0. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

SKILL.md and 23 other files (references) in image of smixs/visual-skills.

  • SKILL.md
  • references/characters.md
  • references/creative-direction.md
  • references/de-slop.md
  • references/dimensional.md
  • references/editing.md
  • references/golden-rules.md
  • references/gpt-image.md
  • references/models.md
  • references/multi-panel.md
  • references/nano-banana.md
  • references/patterns/character-design.md
  • references/patterns/ecommerce.md
  • references/patterns/fashion-editorial.md
  • references/patterns/food-beverage.md
  • references/patterns/portrait-cinema.md
  • references/patterns/poster-illustration.md
  • references/patterns/ui-social.md
  • references/prompt-framework.md
  • … and 5 more

Open the folder on GitHubat commit 92be33a

Compare with similar skills

Image next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Image compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Image this skillsmixs/visual-skills484—~2.1kAutomated safety check: PassCC-BY-4.0
Seedance Storyboard Generatorliangdabiao/Seedance2-Storyboard-Generator2.5k—~2.2kAutomated safety check: PassNone
HiggsfieldOSideMedia/higgsfield-ai-prompt-skill701—~9.1kAutomated safety check: PassMIT
Nbcraftjieyefriic/nbcraft155—~2.8kAutomated safety check: PassMIT
Fal AI Mediamajiayu000/claude-skill-registry6665 repos~1.7kAutomated safety check: PassMIT
Atlas Cloudcalesthio/OpenMontage65k—~1.2kAutomated safety check: PassAGPL-3.0

Similar skills

  • Seedance Storyboard Generator

    liangdabiao/Seedance2-Storyboard-Generator

    专业的Seedance 2.0平台AI视频脚本和分镜生成器。当用户要求:(1) 将文章/故事转换为视频脚本,(2) 生成Seedance 2.0分镜提示词,(3) 规划多集AI视频系列,(4) 为GPT-Image-2、Seedream、Nano Banana…

    2.5k GitHub stars~2.2k tokensUpdated 17 days ago
    Media & CreativeAuto-check passed
  • Higgsfield

    OSideMedia/higgsfield-ai-prompt-skill

    A skill your agent uses whenever the user asks anything about Higgsfield AI — writing or refining video/image prompts, choosing a model (Kling, Veo, Wan, Seedance, Minimax Hailuo, DoP, Soul, Nano…

    701 GitHub stars~9.1k tokensUpdated 11 days ago
    Media & CreativeAuto-check passed
  • Nbcraft

    jieyefriic/nbcraft

    Multi-backend Image + Video Generation CLI (nb command). An agent skill from jieyefriic/nbcraft.

    155 GitHub stars~2.8k tokensUpdated 5 mo ago
    Media & CreativeAuto-check passed
  • Fal AI Media

    majiayu000/claude-skill-registry

    Unified media generation via fal.ai MCP — image, video, and audio.

    666 GitHub starsUsed in 5 repos~1.7k tokens
    Media & CreativeAuto-check passed
  • Atlas Cloud

    calesthio/OpenMontage

    Generate or edit images and videos through the Atlas Cloud gateway.

    65k GitHub stars~1.2k tokensUpdated 4 days ago
    Media & CreativeAuto-check passed
  • Generate

    Blotato-Inc/blotato-skills

    Generate faceless video content cheaply and consistently. An agent skill from Blotato-Inc/blotato-skills.

    183 GitHub stars~2.3k tokensUpdated 1 mo ago
    Media & CreativeAuto-check: notes

More from smixs/visual-skills

  • Video

    smixs/visual-skills

    A skill your agent uses whenever the user asks to create, improve, audit, or split prompts for AI video generators (Seedance, Kling, Veo, Runway, Luma, Pika, Sora, any image-to-video system).

    484 GitHub stars~2.5k tokensUpdated 22 days ago
    Auto-check passed

Questions about Image

What does Image do?

Image prompting skill for Nano Banana (NBP/NB2) and GPT Image 2.5 (Flare/Sunburst). Image is an agent skill from smixs/visual-skills.5 (Flare/Sunburst).

When should I use Image?

Image fits situations like: Сгенерируй картинку; Промпт для картинки; character sheets; vision analysis.

How do I install Image in Claude Code?

Run `npx skills add smixs/visual-skills --skill image -a claude-code`. Or copy the skill folder (image in smixs/visual-skills) into .claude/skills/image in your project. Claude Code loads it when a task matches its description.

How do I install Image in Codex?

Run `npx skills add smixs/visual-skills --skill image -a codex`. Or copy the skill folder (image in smixs/visual-skills) into .agents/skills/image in your project. Codex loads it when a task matches its description.

Can I use Image in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add smixs/visual-skills --skill image -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/image, .gemini/skills/image, .github/skills/image and .opencode/skills/image in your project.

What does Image need to run?

SKILL.md names no scripts, command-line tools or credentials: Image is instructions for the agent only.

Does Image access the network?

SKILL.md names 4 domains. As links in the text: github.com, t.me, sergeshima.com and aimasters.me. This is read from the text; nothing was executed.

Is Image safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Image use?

Image is published under the CC-BY-4.0 licence (declared in SKILL.md). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Image use?

About 2.1k tokens (SKILL.md is roughly 8.6k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full. Its references folder adds about 56k tokens, read only when the agent opens those files.

What are the alternatives to Image?

Skills that share tags, products or a category with Image: Seedance Storyboard Generator (liangdabiao/Seedance2-Storyboard-Generator, 2.5k stars), Higgsfield (OSideMedia/higgsfield-ai-prompt-skill, 701 stars), Nbcraft (jieyefriic/nbcraft, 155 stars) and Fal AI Media (majiayu000/claude-skill-registry, 666 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Image?

smixs (a GitHub user) maintains it in smixs/visual-skills, which has 484 GitHub stars. The repository holds 2 skills in this directory. The repository was last updated on September 16, 2026.

Source: smixs/visual-skills on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.