Agent skill

Oma Image

by first-fluke in first-fluke/oh-my-agent

Multi-vendor AI image generation with authentication-aware parallel dispatch.

MITAuto-check passedMedia & Creative

Install Oma Image

skills CLI
$ npx skills add first-fluke/oh-my-agent --skill oma-image -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install first-fluke/oh-my-agent oma-image --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/first-fluke/oh-my-agent.git skills-src && mkdir -p .claude/skills && cp -r skills-src/benchmarks/runs/oma/.agents/skills/oma-image .claude/skills/oma-image && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
oma-image
GitHub stars
1.3k
Token cost
~3.7k tokens
SKILL.md length
1,608 words
Files
6
Skills in repo
57
Repo updated
First seen
Licence
MIT

At a glance

Multi-vendor AI image generation with authentication-aware parallel dispatch.

  • Works in 3 steps: Validate that the request contains… → Detect attached/reference images and… → Check authentication, cost guardrails,…
  • Image generation
  • SKILL.md covers Scheduling, Structural Flow, Logical Operations and References
  • Calls codex and gemini; needs POLLINATIONS_API_KEY and GEMINI_API_KEY

What it does

Oma Image is an agent skill from first-fluke/oh-my-agent. Multi-vendor AI image generation with authentication-aware parallel dispatch. Routes to Codex (gpt-image-2 via ChatGPT OAuth) and Pollinations (flux/zimage, free with signup). Gemini provider is present but disabled by default (requires billing). Use for image generation, image creation, visual asset generation, and AI art.

Its SKILL.md is about 3.7k tokens, which your agent loads only when the skill is triggered. The skill folder holds 7 other files (for example `config/image-config.yaml`, `resources/checklist.md` and `resources/execution-protocol.md`).

It sits in Media & Creative, covering Image generation. It works with OpenAI. The repository describes itself as: Mechanical verification for AI coding agents — skills pack or full harness (stop-hook gates, artifact checks, independent judges). The licence is MIT.

When your agent uses it

  • Image generation
  • Visual asset generation

Example prompts

  • “/oma-image”

Requirements

  • A credential in POLLINATIONS_API_KEY
  • A credential in GEMINI_API_KEY

Workflow steps

3 steps, taken from the first numbered list in SKILL.md.

  1. Validate that the request contains enough subject, setting, style, usage, and aspect-ratio signal.
  2. Detect attached/reference images and vendor support.
  3. Check authentication, cost guardrails, output path, and count limits.

What it can do on your machine

Read from SKILL.md and the folder at commit b364119. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Shell commands in SKILL.md call:

    • codex
    • gemini

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    Links to these hosts (documentation or services it may open):

    • enter.pollinations.ai

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names these keys or tokens, usually read from environment variables:

    • POLLINATIONS_API_KEY
    • GEMINI_API_KEY

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Oma Image loads about 3.7k tokens when it runs. Until then it costs about 84 tokens; SKILL.md has 1,608 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~84
When it runs · the whole SKILL.md, loaded when a task matches
~3.7k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from first-fluke/oh-my-agent at commit b364119, republished under its MIT licence (© first-fluke). 1,608 words, ~3,699 tokens.

Download SKILL.mdSave it as .claude/skills/oma-image/SKILL.md (or your agent's skills folder). This skill also uses 5 other files; get the full folder from GitHub.
name
oma-image
description
Multi-vendor AI image generation with authentication-aware parallel dispatch. Routes to Codex (gpt-image-2 via ChatGPT OAuth) and Pollinations (flux/zimage, free with signup). Gemini provider is present but disabled by default (requires billing). Use for image generation, image creation, visual asset generation, and AI art.

Image Agent - Multi-Vendor Image Router

Scheduling

Goal

Generate images and visual assets through authenticated multi-vendor routing while preserving prompt clarity, reference-image handling, cost controls, and reproducible output manifests.

Intent signature
  • User asks to generate images, visual assets, illustrations, product photos, concept art, mockups, or AI art.
  • Another skill needs shared image-generation infrastructure.
  • User provides reference images or asks for vendor comparison.
When to use
  • Generating images, visual assets, illustrations, product photos, concept art
  • Comparing output between multiple image models for the same prompt
  • Producing images from prompts within editor workflows (Claude Code, Codex, Gemini CLI)
  • Other skills needing image generation infrastructure (shared invocation)
When NOT to use
  • Editing an existing image or photo manipulation -> out of scope
  • Generating videos or audio -> out of scope
  • Inline vector art / SVG composition from structured data -> use a templating skill
  • Simple asset resizing or format conversion -> use a dedicated image library
Expected inputs
  • Image prompt or creative brief
  • Optional vendor, size, quality, count, output directory, and reference images
  • Authentication/environment state for Codex, Pollinations, or Gemini
Expected outputs
  • Generated image files under .agents/results/images/ or requested output directory
  • manifest.json with prompt, vendor, model, and reproducibility metadata
  • Vendor comparison outputs when --vendor all is used
Dependencies
  • oma image generate CLI and vendor authentication
  • Codex image generation, Pollinations API, or Gemini API/CLI strategy
  • resources/vendor-matrix.md, resources/prompt-tips.md, and config/image-config.yaml
Control-flow features
  • Branches by prompt ambiguity, vendor auth, cost threshold, reference-image support, path safety, and safety/timeout exit codes
  • Calls external vendor APIs/CLIs
  • Reads reference images and writes generated images plus manifests

Structural Flow

Entry
  1. Validate that the request contains enough subject, setting, style, usage, and aspect-ratio signal.
  2. Detect attached/reference images and vendor support.
  3. Check authentication, cost guardrails, output path, and count limits.
Scenes
  1. PREPARE: Clarify or amplify prompt and choose vendor strategy.
  2. ACQUIRE: Validate auth, references, output path, and provider availability.
  3. ACT: Invoke oma image generate with selected vendor(s), prompt, references, and options.
  4. VERIFY: Check manifest, output files, exit code, and provider result.
  5. FINALIZE: Return output paths and relevant warnings.
Transitions
  • If prompt lacks required signal, clarify or show amplified prompt before generation.
  • If --vendor all is requested, require every requested vendor to be available.
  • If reference path is supported by selected vendor, pass it automatically.
  • If estimated cost exceeds guardrail, require confirmation unless bypassed.
Failure and recovery
  • If auth is missing, report vendor-specific authentication requirement.
  • If reference support is unavailable for the selected vendor, reject with actionable guidance.
  • If local CLI is outdated, ask user to run oma update.
  • If generation times out or is blocked, surface exit code and provider status.
Exit
  • Success: images and manifest exist in the output directory.
  • Partial success: some vendors fail in comparison mode and failures are reported.
  • Failure: no image is produced and the route/cost/auth/safety blocker is explicit.

Logical Operations

Actions
ActionSSL primitiveEvidence
Validate prompt completenessVALIDATEClarification protocol
Select vendor strategySELECTVendor matrix and auth state
Read reference imagesREAD--reference paths
Call generation CLI/APICALL_TOOLoma image generate
Write image outputsWRITEImage files and manifest
Validate resultVALIDATEExit code, manifest, files
Report outputNOTIFYFinal path summary
Tools and instruments
  • oma image generate, oma image doctor, oma image list-vendors
  • Codex, Pollinations, and Gemini provider paths
  • Prompt tips, vendor matrix, and image config
Canonical command path
bash
oma image doctor
oma image generate "<prompt>" --vendor auto --size auto --quality auto --format json

With reference images:

bash
oma image generate --reference "<absolute-path>" --vendor codex "<prompt>"
Resource scope
ScopeResource target
LOCAL_FSReference images, generated images, manifests
PROCESSProvider CLIs and image router commands
NETWORKPollinations/Gemini or provider APIs
CREDENTIALSProvider auth and API keys
Preconditions
  • Prompt is sufficiently specified or user approves amplification.
  • Required vendor auth and output permissions exist.
  • Reference paths are accessible when used.
Effects and side effects
  • Creates image files and manifests.
  • May call paid or rate-limited provider APIs.
  • May read attached/reference images.
Guardrails
  1. Clarify before invoking — if the user's request is ambiguous about subject, style, composition, or usage context, ask the user first or amplify the prompt explicitly (showing the user the expanded version for approval). Do NOT silently generate from a vague prompt. See Clarification Protocol below.
  2. Authentication-aware dispatch — detect which vendor CLIs are authenticated and run only those; with --vendor all, every requested vendor must be available (strict).
  3. Cost guardrail — confirm before executing runs whose estimated cost is ≥ $0.20 (configurable). --yes / OMA_IMAGE_YES=1 bypass. Default vendor pollinations (flux/zimage) is free, so auto-triggering on keywords is safe.
  4. Path safety — output paths outside $PWD require --allow-external-out.
  5. Cancellable — SIGINT/SIGTERM aborts in-flight provider calls and the orchestrator.
  6. Deterministic outputs — every run writes manifest.json next to the images for reproducibility.
  7. Max n = 5 — wall-time bound.
  8. Exit codes align with oma search fetch (0, 1, 2=safety, 3=not-found, 4=invalid-input, 5=auth-required, 6=timeout).
Clarification Protocol

Before invoking oma image generate, the calling agent runs this checklist against the user's request. If any answer is "no / unknown", clarify with the user first.

Required signal (must be present or inferable):

  • Subject — what is the primary thing in the image? (object, person, scene)
  • Setting / backdrop — where is it? (context, environment)

Strongly recommended (ask if absent AND not inferable from context):

  • Style — photorealistic, illustration, 3D render, oil painting, concept art, flat vector, …?
  • Mood / lighting — bright vs moody, warm vs cool, dramatic vs minimal
  • Usage context — hero image, icon, thumbnail, product shot, poster? (dictates aspect ratio + composition)
  • Aspect ratio — square (1024x1024), portrait (1024x1536), landscape (1536x1024)?

Amplification shortcut. For brief prompts (e.g. "a red apple"), do not pop clarifying questions if the request is genuinely that simple — instead amplify inline and show the user the expanded version before invoking:

User: "a red apple" Agent: "I'll generate this as: a single glossy red apple centered on a clean white background, soft studio lighting, photorealistic, shallow depth of field, 1024×1024. Shall I proceed, or would you like a different style/composition?"

Skip both clarification and amplification when the user has clearly authored a full creative brief (≥ 2 of: subject + style + lighting + composition). Respect their prompt verbatim.

Category-specific briefs (app mockup, poster, thumbnail, infographic, comic panel, avatar): consult resources/prompt-tips.md → External Prompt Libraries.

Output language. Generation prompts are sent to the provider in English (image models are trained predominantly on English captions). Translate the user's request if they wrote in another language, and show them the translated version during amplification so they can correct misreadings.

Show full SKILL.md (575 more words)Show less
Vendors

This skill follows oh-my-agent's CLI-first concept: whenever a vendor's native CLI can drive generation (and return raw bytes), the subprocess path is preferred over direct API keys. Direct API is only used as a fallback for vendors whose CLI can't yet emit raw image bytes.

VendorStrategyModelsTrigger
codexCLI-first — codex exec via ChatGPT OAuth (codex login), built-in image_gengpt-image-2Logged in via Codex CLI (no API key)
pollinationsDirect HTTP — gen.pollinations.ai/v1/images/generations (free signup for key)Free: flux, zimage. Credit-gated: qwen-image, wan-image, gpt-image-2, klein, kontext, gptimage, gptimage-largePOLLINATIONS_API_KEY set (free at https://enter.pollinations.ai). No native CLI exists.
geminiCLI-first fallback → direct API. gemini -p (stream) is the preferred path but currently disabled at precheck (CLI's agentic loop does not return raw inlineData bytes on stdout as of Gemini CLI 0.38). Until the CLI exposes a non-agentic image surface, the provider falls back to the direct generativelanguage.googleapis.com API.gemini-2.5-flash-image, gemini-3.1-flash-image-previewPreferred: gemini auth login. Fallback: GEMINI_API_KEY + billing.
Invocation
Standalone
/oma-image a red apple on white background
/oma-image --vendor all --size 1536x1024 jeju coastline at sunset
/oma-image -n 3 --quality high --out ./hero "minimalist dashboard hero illustration"
Shell CLI
oma image generate "<prompt>" [--vendor auto|codex|pollinations|gemini|all] [-n 1..5] \
                             [--size 1024x1024|1024x1536|1536x1024|auto] \
                             [--quality low|medium|high|auto] \
                             [--out <dir>] [--allow-external-out] \
                             [-r <path>]... \
                             [--timeout 180] [-y] [--no-prompt-in-manifest] \
                             [--dry-run] [--format text|json]
oma image doctor
oma image list-vendors

Gemini-only escalation flag: --strategy mcp,stream,api (overrides vendors.gemini.strategies).

Reference Images (-r, --reference)

Attach up to 10 reference images (PNG/JPEG/GIF/WebP, ≤ 5MB each) to guide style, subject identity, or composition. Repeatable or comma-separated.

oma image generate -r ~/Downloads/otter.jpeg "same otter in dramatic lighting"
oma image generate -r a.png -r b.png "blend these two styles"

Supported vendors:

VendorSupportHow
codex (gpt-image-2)✅Passes -i <path> to codex exec
gemini (2.5-flash-image)✅Inlines base64 inlineData parts in request
pollinations❌Rejected with exit code 4 (requires URL hosting; see PR #2 roadmap)

Paths: absolute or relative to $CWD. Host CLIs usually expose attached images via:

  • Claude Code: ~/.claude/image-cache/<session>/N.png (surfaced in system messages as [Image: source: <path>])
  • Antigravity: workspace upload directory (exact path shown in IDE)
  • Codex CLI as host: user must pass the filesystem path explicitly; in-conversation attachments are not forwarded
Agent Behavior: Auto-forward Attached References (MANDATORY)

When ALL of the following are true, the calling agent MUST pass the attached image via --reference <path> automatically. Never describe the image in prose as a workaround.

  1. The user asks to generate or edit an image (referencing the attached one by phrases like "이거", "this image", "same style as this", "이 수달", etc.).
  2. A host-surfaced attached image is visible to the agent — e.g. a Claude Code system message with [Image: source: <path>], or an Antigravity workspace upload path, or an explicit filesystem path in the user's message.
  3. The selected vendor supports references (codex or gemini).

Required action: invoke oma image generate --reference <absolute-path> --vendor <codex|gemini> "<prompt>". If the user didn't specify a vendor, default to codex (CLI-first, widest availability). Do NOT:

  • Fall back to prose description ("I'll describe the otter's appearance...").
  • Ask the user to re-type or re-attach the path.
  • Claim the CLI doesn't support references without first running oma image generate --help to verify.

If the local CLI is outdated (--reference is missing from --help): tell the user to run oma update once, then retry. Do not silently degrade to prose.

If the reference path is from Claude Code's image-cache: note to the user that the path is session-scoped and suggest copying the file to a durable location if they want to reuse it later. Still proceed with the generation.

Shared Infrastructure (from other skills)

Other skills call oma image generate --format json and parse the JSON manifest from stdout.

Output Layout
.agents/results/images/
├── 20260424-143052-ab12cd/                    # single-vendor run
│   └── pollinations-flux.jpg
│       (or codex-gpt-image-2.png)
│       manifest.json
└── 20260424-143122-7z9kqw-compare/            # --vendor all run
    ├── codex-gpt-image-2.png
    ├── pollinations-flux.jpg
    └── manifest.json

References

Follow resources/execution-protocol.md step by step. See resources/vendor-matrix.md for strategy precheck rules. Use resources/prompt-tips.md for writing effective prompts. Before submitting, run resources/checklist.md.

Configuration

Project-specific settings: config/image-config.yaml. Env vars: OMA_IMAGE_DEFAULT_VENDOR, OMA_IMAGE_DEFAULT_OUT, OMA_IMAGE_YES, POLLINATIONS_API_KEY, GEMINI_API_KEY, OMA_IMAGE_GEMINI_STRATEGIES.

  • Execution steps: resources/execution-protocol.md
  • Vendor matrix: resources/vendor-matrix.md
  • Prompt tips: resources/prompt-tips.md
  • Checklist: resources/checklist.md
  • Context loading: ../_shared/core/context-loading.md

© first-fluke, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

SKILL.md and 5 other files in benchmarks/runs/oma/.agents/skills/oma-image of first-fluke/oh-my-agent.

  • SKILL.md
  • config/image-config.yaml
  • resources/checklist.md
  • resources/execution-protocol.md
  • resources/prompt-tips.md
  • resources/vendor-matrix.md

Open the folder on GitHubat commit b364119

Compare with similar skills

Oma Image next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Oma Image compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Oma Image this skillfirst-fluke/oh-my-agent1.3k—~3.7kAutomated safety check: PassMIT
AI Image Generation and Editingzhayujie/CowAgent47k—~1.3kAutomated safety check: PassMIT
GPT Image Generation CLIwuyoscar/GPT-Image2-Skill5.7k—~2.5kAutomated safety check: NotesMIT
Imagegentheowenyoung/home1154 repos~4.8kAutomated safety check: PassApache-2.0
Openai Image Gentrpc-group/trpc-agent-go1.9k12 repos~843Automated safety check: PassApache-2.0
Image Generationonyx-dot-app/onyx32k1 repos~1.7kAutomated safety check: PassCustom licence

Similar skills

  • Generates or edits images from text prompts through a Python script that picks an image backend based on which API keys are configured.

    47k GitHub stars~1.3k tokensUpdated today
    Media & CreativeAuto-check passed
  • GPT Image Generation CLI

    wuyoscar/GPT-Image2-Skill

    Generates and edits images with GPT Image 2 or 2.5 through a packaged CLI and a prompt gallery, after settling which model fits the request.

    5.7k GitHub stars~2.5k tokensUpdated 8 days ago
    Media & CreativeAuto-check: notes
  • Imagegen

    theowenyoung/home

    Generate or edit raster images when the task benefits from AI-created bitmap visuals such as photos, illustrations, textures, sprites, mockups, or transparent-background cutouts.

    115 GitHub starsUsed in 4 repos~4.8k tokens
    Media & CreativeAuto-check passed
  • Openai Image Gen

    trpc-group/trpc-agent-go

    Batch-generate images via OpenAI Images API. An agent skill from trpc-group/trpc-agent-go.

    1.9k GitHub starsUsed in 12 repos~843 tokens
    Media & CreativeAuto-check passed
  • Image Generation

    onyx-dot-app/onyx

    Generate or edit raster images (photos, illustrations, textures, sprites, mockups, logos, infographics) using the workspace's configured image-generation provider via onyx-cli image.

    32k GitHub starsUsed in 1 repo~1.7k tokens
    Media & CreativeAuto-check passed
  • BlockRun Image Generation

    BlockRunAI/ClawRouter

    Generates or edits images through ClawRouter's local image API, with a choice of models and sizes and payment handled automatically through x402.

    6.6k GitHub stars~2.1k tokensUpdated 3 days ago
    Media & CreativeAuto-check passed

More from first-fluke/oh-my-agent

All 57 skills in this repo
  • OMA Multi-Agent Orchestration

    first-fluke/oh-my-agent

    Decomposes a complex feature into tasks, dispatches parallel specialist agents with durable state, and supervises verification, QA review and retries.

    1.3k GitHub stars~4.1k tokensUpdated today
    Auto-check passed
  • OMA Multi-Agent Orchestrator

    first-fluke/oh-my-agent

    Splits a complex feature into prioritized tasks, spawns specialist CLI subagents in parallel, tracks them through shared memory and verifies each result.

    1.3k GitHub stars~3.1k tokensUpdated today
    Auto-check passed
  • Architecture Decisions and ADRs

    first-fluke/oh-my-agent

    Evaluates system boundaries and tradeoffs and writes architecture recommendations, option comparisons or ADRs, with a Mermaid diagram when structure changes.

    1.3k GitHub stars~2.6k tokensUpdated today
    Auto-check passed
  • OMA Backend Agent

    first-fluke/oh-my-agent

    Backend specialist for APIs, database work, authentication and migrations that follows clean architecture with router, service and repository layers.

    1.3k GitHub stars~2.4k tokensUpdated today
    Auto-check passed
  • oma Bootstrap

    first-fluke/oh-my-agent

    Installs or checks the oma CLI and its runtimes (bun, uv, serena) in a fresh workspace so that oma-* skills can run their commands.

    1.3k GitHub stars~719 tokensUpdated today
    Auto-check passed
  • Design-First Brainstorm

    first-fluke/oh-my-agent

    Explores intent, constraints and alternative approaches before any planning, working through questions one at a time and saving an approved design for later steps.

    1.3k GitHub stars~1.7k tokensUpdated today
    Auto-check passed

Works with

Questions about Oma Image

What does Oma Image do?

Multi-vendor AI image generation with authentication-aware parallel dispatch. Oma Image is an agent skill from first-fluke/oh-my-agent. Multi-vendor AI image generation with authentication-aware parallel dispatch.

When should I use Oma Image?

Oma Image fits situations like: image generation; visual asset generation.

How do I install Oma Image in Claude Code?

Run `npx skills add first-fluke/oh-my-agent --skill oma-image -a claude-code`. Or copy the skill folder (benchmarks/runs/oma/.agents/skills/oma-image in first-fluke/oh-my-agent) into .claude/skills/oma-image in your project. Claude Code loads it when a task matches its description.

How do I install Oma Image in Codex?

Run `npx skills add first-fluke/oh-my-agent --skill oma-image -a codex`. Or copy the skill folder (benchmarks/runs/oma/.agents/skills/oma-image in first-fluke/oh-my-agent) into .agents/skills/oma-image in your project. Codex loads it when a task matches its description.

Can I use Oma Image in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add first-fluke/oh-my-agent --skill oma-image -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/oma-image, .gemini/skills/oma-image, .github/skills/oma-image and .opencode/skills/oma-image in your project.

What does Oma Image need to run?

Going by SKILL.md and its folder, Oma Image needs the command-line tools its instructions call (codex and gemini) and credentials named POLLINATIONS_API_KEY and GEMINI_API_KEY. Our summary lists: A credential in POLLINATIONS_API_KEY; A credential in GEMINI_API_KEY.

Does Oma Image access the network?

SKILL.md names 1 domain. As links in the text: enter.pollinations.ai. This is read from the text; nothing was executed.

Is Oma Image safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Oma Image use?

Oma Image is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Oma Image use?

About 3.7k tokens (SKILL.md is roughly 15k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Oma Image?

Skills that share tags, products or a category with Oma Image: AI Image Generation and Editing (zhayujie/CowAgent, 47k stars), GPT Image Generation CLI (wuyoscar/GPT-Image2-Skill, 5.7k stars), Imagegen (theowenyoung/home, 115 stars) and Openai Image Gen (trpc-group/trpc-agent-go, 1.9k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Oma Image?

first-fluke (a GitHub organization) maintains it in first-fluke/oh-my-agent, which has 1,338 GitHub stars. The repository holds 57 skills in this directory. The repository was last updated on October 9, 2026.

Source: first-fluke/oh-my-agent on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.