Generate/edit images via fal.ai. An agent skill from artwist-polyakov/polyakov-claude-skills.

MITAuto-check: notesMedia & Creative

Install Fal AI Image

skills CLI
$ npx skills add artwist-polyakov/polyakov-claude-skills --skill fal-ai-image -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install artwist-polyakov/polyakov-claude-skills fal-ai-image --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/artwist-polyakov/polyakov-claude-skills.git skills-src && mkdir -p .claude/skills && cp -r skills-src/plugins/fal-ai-image/skills/fal-ai-image .claude/skills/fal-ai-image && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
fal-ai-image
GitHub stars
208
Token cost
~2.1k tokens
SKILL.md length
794 words
Files
14 (incl. scripts, references)
Skills in repo
21
Repo updated
First seen
Licence
MIT

At a glance

Generate/edit images via fal.ai. An agent skill from artwist-polyakov/polyakov-claude-skills.

  • Works in 4 steps: model — one-off override for the current… → FAL_IMAGE_MODEL — exact override in config → FAL_IMAGE_PROVIDER — google or openai → …
  • Tasks that involve Image generation
  • SKILL.md covers STOP — Read Before Acting, Quick Start Decision, Config and Model Notes, plus 4 more sections
  • Runs Shell scripts from its folder; calls sh; needs FAL_KEY

What it does

Fal AI Image is an agent skill from artwist-polyakov/polyakov-claude-skills. Generate/edit images via fal.ai. Supports Google Nano Banana Pro and OpenAI GPT Image 2 selected from config/.env. Supports reference images and strong text rendering. ALWAYS read SKILL.md before first use.

Its SKILL.md is about 2.1k tokens, which your agent loads only when the skill is triggered. The skill folder holds 17 other files, including scripts and reference files (for example `config/README.md`, `references/EDITING.md` and `references/MODELS.md`).

It sits in Media & Creative, covering Image generation. It works with fal, Google Gemini and OpenAI. The repository describes itself as: Набор скиллов для Claude. The licence is MIT.

When your agent uses it

  • Tasks that involve Image generation

Example prompts

  • “/fal-ai-image”

Requirements

  • A Bash shell
  • A credential in FAL_KEY

Workflow steps

4 steps, taken from the first numbered list in SKILL.md.

  1. model — one-off override for the current command
  2. FAL_IMAGE_MODEL — exact override in config
  3. FAL_IMAGE_PROVIDER — google or openai
  4. nothing set — default to Nano Banana Pro

What it can do on your machine

Read from SKILL.md and the folder at commit 8bbeead. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Ships 8 files in scripts/ (Shell), which the agent can run.

    Shell commands in SKILL.md call:

    • sh

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names these keys or tokens, usually read from environment variables:

    • FAL_KEY

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Fal AI Image loads about 2.1k tokens when it runs, and up to ~3.1k if it reads all its reference files. Until then it costs about 55 tokens; SKILL.md has 794 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~55
When it runs · the whole SKILL.md, loaded when a task matches
~2.1k
With references · SKILL.md plus every file in references/, read only if the agent opens them
~3.1k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check: notes

The automated check noted patterns worth knowing about, such as sudo or a known installer.

  • NoteMentions a .env fileSKILL.md:3
    OpenAI GPT Image 2 selected from config/.env. Supports reference images and strong text rendering. ALWAYS read SKILL.md
  • NoteMentions a .env fileSKILL.md:11
    nai/gpt-image-2`) — enabled from `config/.env`
  • NoteMentions a .env fileSKILL.md:35
    Model comes from config/.env:
  • NoteMentions a .env fileSKILL.md:43
    Requires `FAL_KEY` in `config/.env` or the environment.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); the scripts in this folder are not scanned.

SKILL.md

The full file from artwist-polyakov/polyakov-claude-skills at commit 8bbeead, republished under its MIT licence (© artwist-polyakov). 794 words, ~2,081 tokens.

Download SKILL.mdSave it as .claude/skills/fal-ai-image/SKILL.md (or your agent's skills folder). This skill also uses 13 other files; get the full folder from GitHub.
name
fal-ai-image
description
Generate/edit images via fal.ai. Supports Google Nano Banana Pro and OpenAI GPT Image 2 selected from config/.env. Supports reference images and strong text rendering. ALWAYS read SKILL.md before first use.

fal-ai-image

Generate images via fal.ai. The skill now supports two fal-hosted models:

  • Google Nano Banana Pro (fal-ai/nano-banana-pro) — default, backward-compatible
  • OpenAI GPT Image 2 (openai/gpt-image-2) — enabled from config/.env

Synonyms the agent should treat as equivalent:

  • gpt = openai = GPT Image 2
  • nano banana = google = gemini = Nano Banana Pro

Best for: infographics, text rendering, banners, photo edits, reference-based compositions.

STOP — Read Before Acting

  • DO NOT use Pillow, ImageMagick, or post-processing for text/logo overlay unless the user explicitly asked for that workflow
  • DO NOT use Generate mode when the user provided reference images — use Edit mode
  • DO NOT assume GPT is active — check config/README.md logic; if no selector is configured, the skill stays on Nano Banana
  • DO NOT guess provider-specific params — Nano Banana and GPT Image 2 use different schemas
  • DO pass --model ... when the user explicitly asks for a specific model/provider or asks to compare providers
  • DO NOT skip uploading local files — run upload.sh first to get URLs for edit.sh

Quick Start Decision

text
Reference images provided?  -> Edit mode   (upload.sh -> edit.sh)
Text-only generation?       -> Generate mode (generate.sh)

Model comes from config/.env:
  no selector set           -> Nano Banana Pro
  FAL_IMAGE_PROVIDER=openai -> GPT Image 2
  FAL_IMAGE_MODEL=...       -> exact override

Config

Requires FAL_KEY in config/.env or the environment.

Model selection:

  1. --model — one-off override for the current command
  2. FAL_IMAGE_MODEL — exact override in config
  3. FAL_IMAGE_PROVIDER — google or openai
  4. nothing set — default to Nano Banana Pro

Use --model whenever the user explicitly says things like:

  • "сделай через GPT"
  • "используй OpenAI"
  • "сделай через Google / Gemini / Nano Banana"
  • "какие тут есть провайдеры?"

Answer that the skill supports two choices and map the command like this:

  • --model gpt for OpenAI GPT Image 2
  • --model gemini or --model nano-banana for Nano Banana

OpenAI quality default:

  • FAL_IMAGE_OPENAI_QUALITY=medium unless overridden with --quality
  • this skill intentionally uses medium as the default for GPT to avoid expensive exploratory runs

Full setup and troubleshooting: config/README.md.

Read references only when needed:

Model Notes

Nano Banana Pro

Strengths:

  • lower-friction default for existing installs
  • strong text rendering, including Cyrillic
  • good at infographics, banners, and mixed text/image layouts
  • supports --web-search in generate mode

Main params:

  • --aspect-ratio
  • --resolution
  • --web-search (generate only)
OpenAI GPT Image 2

Strengths:

  • stronger prompt adherence and fine-grained edits
  • better photorealism and product-style renders
  • native quality control
  • uses the same FAL_KEY through fal, no separate OpenAI key
  • current fal pricing is size/quality-dependent; a rough high-quality mental model is about $0.18 per image, but check the live model page before quoting an exact number

Main params:

  • --image-size
  • --quality

Compatibility layer:

  • if --image-size is omitted, the scripts derive a valid OpenAI image_size from --aspect-ratio + --resolution
  • this lets old prompts continue working after the provider switch in config

Workflow

Generate mode
  1. Decide model from config
  2. Clarify missing params only if needed:
    • Nano Banana: aspect ratio, resolution
    • GPT Image 2: image size and quality
  3. Propose save path based on project structure
  4. Run generate.sh
  5. Parse result JSON, report URL and local files if downloaded
Show full SKILL.md (327 more words)Show less
Edit mode
  1. Get reference images:
    • URL already available -> use directly
    • local file -> upload.sh
  2. Decide model from config
  3. Clarify edit intent
  4. Run edit.sh
  5. Parse result JSON, report URL and local files if downloaded

Scripts

generate.sh

Nano Banana example:

bash
sh scripts/generate.sh \
  --model "gemini" \
  --prompt "infographic about coffee brewing" \
  --aspect-ratio "9:16" \
  --resolution "1K" \
  --output-dir "./images" \
  --filename "coffee_infographic"

GPT Image 2 example:

bash
sh scripts/generate.sh \
  --model "gpt" \
  --prompt "realistic product hero shot with sharp packaging text" \
  --image-size "landscape_4_3" \
  --quality "medium" \
  --output-dir "./images" \
  --filename "product_hero"

Compatibility example for GPT:

bash
sh scripts/generate.sh \
  --model "openai" \
  --prompt "editorial portrait, window light, magazine cover layout" \
  --aspect-ratio "4:3" \
  --resolution "2K"
ParamRequiredDefaultNotes
--promptyes-text prompt
--modelnoconfig / Nano Banana fallbacknano-banana, google, gemini, gpt, openai, or exact endpoint
--aspect-rationo1:1Nano native; for GPT used only when --image-size is omitted
--resolutionno1KNano native; for GPT used only when --image-size is omitted
--image-sizenoderived from ratio/resolutionGPT only; preset (landscape_4_3) or WIDTHxHEIGHT
--qualitynomedium via configGPT only; low, medium, high
--num-imagesno11-4
--output-formatnopngjpeg, png, webp
--output-dirno-local path
--filenamenogeneratedbase filename
--web-searchnofalseNano only; ignored for GPT
edit.sh

Nano Banana example:

bash
sh scripts/edit.sh \
  --model "gemini" \
  --prompt "combine these into a collage" \
  --image-urls "https://example.com/img1.png,https://example.com/img2.png" \
  --aspect-ratio "16:9" \
  --output-dir "./images" \
  --filename "collage"

GPT Image 2 example:

bash
sh scripts/edit.sh \
  --model "gpt" \
  --prompt "make this product shot look like a premium studio campaign" \
  --image-urls "https://example.com/source.png" \
  --mask-url "https://example.com/mask.png" \
  --image-size "auto" \
  --quality "medium" \
  --output-dir "./images" \
  --filename "studio_edit"
ParamRequiredDefaultNotes
--promptyes-edit instruction
--image-urlsyes-comma-separated URLs
--modelnoconfig / Nano Banana fallbacknano-banana, google, gemini, gpt, openai, or exact endpoint
--mask-urlno-GPT edit only; optional mask for targeted edits
--aspect-rationoautoNano native; for GPT used only when --image-size is omitted
--resolutionno1KNano native; for GPT used only when --image-size is omitted
--image-sizenoderived from ratio/resolution / autoGPT only
--qualitynomedium via configGPT only
--num-imagesno11-4
--output-formatnopngjpeg, png, webp
--output-dirno-local path
--filenamenoeditedbase filename
upload.sh
bash
# Get hosted URL for local file
URL=$(sh scripts/upload.sh --file /path/to/image.png)

# Get base64 data URI for manual API work
URI=$(sh scripts/upload.sh --file /path/to/image.png --base64)

Cost Guidance

  • Nano Banana Pro is the cheaper and safer default for quick iterations
  • GPT Image 2 cost depends heavily on quality and image_size
  • this skill defaults GPT to medium quality to reduce surprise spend

For current pricing, check fal's model pages in config/README.md.

Notes

  • result URLs expire in roughly one hour — download locally if you need persistence
  • uploaded files on fal storage are temporary
  • edit.sh now polls the same /edit queue endpoints documented by fal

© artwist-polyakov, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

SKILL.md and 13 other files (scripts, references) in plugins/fal-ai-image/skills/fal-ai-image of artwist-polyakov/polyakov-claude-skills.

  • SKILL.md
  • .gitignore
  • config/.env.example
  • config/README.md
  • references/EDITING.md
  • references/MODELS.md
  • scripts/common.sh
  • scripts/edit.sh
  • scripts/generate.sh
  • scripts/tests/run.sh
  • scripts/tests/test_config_selector.sh
  • scripts/tests/test_edit_openai_mask.sh
  • scripts/tests/test_openai_image_size.sh
  • scripts/upload.sh

Open the folder on GitHubat commit 8bbeead

Compare with similar skills

Fal AI Image next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Fal AI Image compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Fal AI Image this skillartwist-polyakov/polyakov-claude-skills208—~2.1kAutomated safety check: NotesMIT
Model Routingfal-ai-community/skills251—~1.5kAutomated safety check: PassNone
9Router Image Generationdecolua/9router31k—~830Automated safety check: PassMIT
2D Map and Scene Generator0x0funky/agent-sprite-forge4.4k—~2.9kAutomated safety check: PassMIT
Forge Media Route Layer0x0funky/agent-sprite-forge4.4k—~2.2kAutomated safety check: PassMIT
2D Sprite Generator0x0funky/agent-sprite-forge4.4k—~3.6kAutomated safety check: PassMIT

Similar skills

  • Model Routing

    fal-ai-community/skills

    Choose default fal.ai endpoint IDs for genmedia production skills.

    251 GitHub stars~1.5k tokensUpdated 12 days ago
    Media & CreativeAuto-check passed
  • Generates images through a 9Router gateway's image endpoint, with model discovery, the request fields and per-provider quirks for OpenAI, Gemini, MiniMax and others.

    31k GitHub stars~830 tokensUpdated 3 days ago
    Media & CreativeAuto-check passed
  • 2D Map and Scene Generator

    0x0funky/agent-sprite-forge

    Plans and builds 2D game maps and scenes, from tilemaps and parallax backgrounds to HD-2D plates, with collision checks, a playable HTML preview and Tiled, Godot or LDtk export.

    4.4k GitHub stars~2.9k tokensUpdated 5 days ago
    Game DevelopmentAuto-check passed
  • Forge Media Route Layer

    0x0funky/agent-sprite-forge

    Generates an image or an image-to-video clip through a configured provider API or a signed-in Codex or Grok CLI, and reports the route, file, hash and cost estimate.

    4.4k GitHub stars~2.2k tokensUpdated 5 days ago
    Media & CreativeAuto-check passed
  • 2D Sprite Generator

    0x0funky/agent-sprite-forge

    Produces game-ready 2D characters, creatures, props, icons and effects as master stills, sheets or clips, and exports frames for common game engines.

    4.4k GitHub stars~3.6k tokensUpdated 5 days ago
    Game DevelopmentAuto-check passed
  • Nano Banana Pro Prompts Recommend Skill

    YouMind-OpenLab/nano-banana-pro-prompts-recommend-skill

    Recommend suitable prompts from 10,000+ Nano Banana Pro image generation prompts based on user needs.

    1.9k GitHub starsUsed in 1 repo~4.1k tokens
    Media & CreativeAuto-check passed

More from artwist-polyakov/polyakov-claude-skills

All 21 skills in this repo
  • Agent Deck Sessions

    artwist-polyakov/polyakov-claude-skills

    Launches, monitors and collects results from child AI agent sessions with the agent-deck terminal session manager.

    208 GitHub stars~965 tokensUpdated 3 days ago
    Auto-check passed
  • Crawl4AI SEO Site Crawler

    artwist-polyakov/polyakov-claude-skills

    Crawls a site with Crawl4AI to audit titles, meta tags, H1s, canonicals, navigation and internal links, and to compare landing pages and competitor sites.

    208 GitHub stars~1.7k tokensUpdated 3 days ago
    Auto-check passed
  • GitHub Pages Publisher

    artwist-polyakov/polyakov-claude-skills

    Publishes already-built static pages to a GitHub Pages repo under a year, year-month and slug folder layout, optimizes large images and returns the public URL.

    208 GitHub stars~1.2k tokensUpdated 3 days ago
    Auto-check passed
  • Perplexity Search and Research

    artwist-polyakov/polyakov-claude-skills

    Shell scripts for web search and research through the Perplexity API: raw results, cited answers, background deep research and page fetching, with results cached on disk.

    208 GitHub stars~2.5k tokensUpdated 3 days ago
    Auto-check: notes
  • Sourcecraft Publisher

    artwist-polyakov/polyakov-claude-skills

    Publish static page artifacts to SourceCraft Sites (Yandex infrastructure, works in Russia), with advisory image optimization and an original-image path.

    208 GitHub stars~990 tokensUpdated 3 days ago
    Auto-check: notes
  • X/Twitter Research via Grok

    artwist-polyakov/polyakov-claude-skills

    Searches X/Twitter through the xAI Grok API to build digests, trend reports and thread analysis, meant to surface post ideas for a Telegram channel.

    208 GitHub stars~1.8k tokensUpdated 3 days ago
    Auto-check: notes

Questions about Fal AI Image

What does Fal AI Image do?

Generate/edit images via fal.ai. An agent skill from artwist-polyakov/polyakov-claude-skills. Fal AI Image is an agent skill from artwist-polyakov/polyakov-claude-skills.ai.

When should I use Fal AI Image?

Fal AI Image fits situations like: tasks that involve Image generation.

How do I install Fal AI Image in Claude Code?

Run `npx skills add artwist-polyakov/polyakov-claude-skills --skill fal-ai-image -a claude-code`. Or copy the skill folder (plugins/fal-ai-image/skills/fal-ai-image in artwist-polyakov/polyakov-claude-skills) into .claude/skills/fal-ai-image in your project. Claude Code loads it when a task matches its description.

How do I install Fal AI Image in Codex?

Run `npx skills add artwist-polyakov/polyakov-claude-skills --skill fal-ai-image -a codex`. Or copy the skill folder (plugins/fal-ai-image/skills/fal-ai-image in artwist-polyakov/polyakov-claude-skills) into .agents/skills/fal-ai-image in your project. Codex loads it when a task matches its description.

Can I use Fal AI Image in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add artwist-polyakov/polyakov-claude-skills --skill fal-ai-image -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/fal-ai-image, .gemini/skills/fal-ai-image, .github/skills/fal-ai-image and .opencode/skills/fal-ai-image in your project.

What does Fal AI Image need to run?

Going by SKILL.md and its folder, Fal AI Image needs a shell for the scripts in its folder, the command-line tools its instructions call (sh) and credentials named FAL_KEY. Our summary lists: A Bash shell; A credential in FAL_KEY.

Does Fal AI Image access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is Fal AI Image safe to install?

Our automated static check of SKILL.md found notes only (mentions a .env file), nothing it rates as a warning. It is not a guarantee. The check reads SKILL.md only: the scripts in the folder are not scanned, so read them before running anything.

What licence does Fal AI Image use?

Fal AI Image is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Fal AI Image use?

About 2.1k tokens (SKILL.md is roughly 8.3k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full. Its references folder adds about 1k tokens, read only when the agent opens those files.

What are the alternatives to Fal AI Image?

Skills that share tags, products or a category with Fal AI Image: Model Routing (fal-ai-community/skills, 251 stars), 9Router Image Generation (decolua/9router, 31k stars), 2D Map and Scene Generator (0x0funky/agent-sprite-forge, 4.4k stars) and Forge Media Route Layer (0x0funky/agent-sprite-forge, 4.4k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Fal AI Image?

artwist-polyakov (a GitHub user) maintains it in artwist-polyakov/polyakov-claude-skills, which has 208 GitHub stars. The repository holds 21 skills in this directory. The repository was last updated on October 8, 2026.

Source: artwist-polyakov/polyakov-claude-skills on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.