Imagen
sanjay3290/ai-skills
Generate images using Google Gemini's image generation capabilities.
Production prompt director for GPT Image 2 (imagegen20). An agent skill from alecs5am/ralphy.
$ npx skills add alecs5am/ralphy --skill gpt-image-2-director -a claude-codeProject install by default; add -g for ~/.claude/skills/.
$ gh skill install alecs5am/ralphy gpt-image-2-director --agent claude-codeProject scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).
$ git clone --depth 1 https://github.com/alecs5am/ralphy.git skills-src && mkdir -p .claude/skills && cp -r skills-src/notes/skills/gpt-image-2-director .claude/skills/gpt-image-2-director && rm -rf skills-srcUse ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.
Claude Code skills documentation · loads skills from .claude/skills/
Install the "gpt-image-2-director" agent skill from https://github.com/alecs5am/ralphy/tree/main/notes/skills/gpt-image-2-director into .claude/skills/gpt-image-2-director/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "gpt-image-2-director", then confirm the skill loads.Claude Code copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$skill-installer install https://github.com/alecs5am/ralphy/tree/main/notes/skills/gpt-image-2-directorType this inside Codex. $skill-installer <name> installs a curated skill from openai/skills. The installer writes to $CODEX_HOME/skills (default ~/.codex/skills). Restart Codex if the skill does not show up.
$ npx skills add alecs5am/ralphy --skill gpt-image-2-director -a codexProject install goes to .agents/skills/; add -g for ~/.codex/skills/.
$ gh skill install alecs5am/ralphy gpt-image-2-director --agent codexProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/alecs5am/ralphy.git skills-src && mkdir -p .agents/skills && cp -r skills-src/notes/skills/gpt-image-2-director .agents/skills/gpt-image-2-director && rm -rf skills-srcUse ~/.agents/skills/ instead of .agents/skills for a personal install.
Codex skills documentation · loads skills from .agents/skills/
Install the "gpt-image-2-director" agent skill from https://github.com/alecs5am/ralphy/tree/main/notes/skills/gpt-image-2-director into .agents/skills/gpt-image-2-director/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "gpt-image-2-director", then confirm the skill loads.Codex copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add alecs5am/ralphy --skill gpt-image-2-director -a cursorProject install goes to .agents/skills/; add -g for ~/.cursor/skills/.
$ gh skill install alecs5am/ralphy gpt-image-2-director --agent cursorProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/alecs5am/ralphy.git skills-src && mkdir -p .cursor/skills && cp -r skills-src/notes/skills/gpt-image-2-director .cursor/skills/gpt-image-2-director && rm -rf skills-srcUse ~/.cursor/skills/ instead of .cursor/skills for a personal install.
Cursor skills documentation · loads skills from .cursor/skills/, .agents/skills/, .claude/skills/, .codex/skills/
Install the "gpt-image-2-director" agent skill from https://github.com/alecs5am/ralphy/tree/main/notes/skills/gpt-image-2-director into .cursor/skills/gpt-image-2-director/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "gpt-image-2-director", then confirm the skill loads.Cursor copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gemini skills install https://github.com/alecs5am/ralphy.git --path notes/skills/gpt-image-2-director--scope user (default) or --scope workspace; --path is the subfolder of the repo that holds the skill; --consent skips the security confirmation prompt.
$ npx skills add alecs5am/ralphy --skill gpt-image-2-director -a gemini-cliProject install goes to .agents/skills/; add -g for ~/.gemini/skills/.
$ gh skill install alecs5am/ralphy gpt-image-2-director --agent gemini-cliProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/alecs5am/ralphy.git skills-src && mkdir -p .gemini/skills && cp -r skills-src/notes/skills/gpt-image-2-director .gemini/skills/gpt-image-2-director && rm -rf skills-srcUse ~/.gemini/skills/ instead of .gemini/skills for a personal install, then run /skills reload.
Gemini CLI skills documentation · loads skills from .gemini/skills/, .agents/skills/
Install the "gpt-image-2-director" agent skill from https://github.com/alecs5am/ralphy/tree/main/notes/skills/gpt-image-2-director into .gemini/skills/gpt-image-2-director/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "gpt-image-2-director", then confirm the skill loads.Gemini CLI copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gh skill install alecs5am/ralphy gpt-image-2-directorInstalls for Copilot at project scope by default; add --scope user for a personal install. Preview a skill first with gh skill preview. Needs GitHub CLI 2.90.0 or later (public preview).
$ npx skills add alecs5am/ralphy --skill gpt-image-2-director -a github-copilotProject install goes to .agents/skills/; add -g for ~/.copilot/skills/.
$ git clone --depth 1 https://github.com/alecs5am/ralphy.git skills-src && mkdir -p .github/skills && cp -r skills-src/notes/skills/gpt-image-2-director .github/skills/gpt-image-2-director && rm -rf skills-srcUse ~/.copilot/skills/ instead of .github/skills for a personal install. Commit .github/skills so cloud agent and code review can use it.
GitHub Copilot skills documentation · loads skills from .github/skills/, .claude/skills/, .agents/skills/
Install the "gpt-image-2-director" agent skill from https://github.com/alecs5am/ralphy/tree/main/notes/skills/gpt-image-2-director into .github/skills/gpt-image-2-director/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "gpt-image-2-director", then confirm the skill loads.GitHub Copilot copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add alecs5am/ralphy --skill gpt-image-2-director -a opencodeOpenCode documents no install command of its own. Project install goes to .agents/skills/; add -g for ~/.config/opencode/skills/.
$ gh skill install alecs5am/ralphy gpt-image-2-director --agent opencodeProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/alecs5am/ralphy.git skills-src && mkdir -p .opencode/skills && cp -r skills-src/notes/skills/gpt-image-2-director .opencode/skills/gpt-image-2-director && rm -rf skills-srcUse ~/.config/opencode/skills/ instead of .opencode/skills for a personal install.
OpenCode skills documentation · loads skills from .opencode/skills/, .claude/skills/, .agents/skills/
Install the "gpt-image-2-director" agent skill from https://github.com/alecs5am/ralphy/tree/main/notes/skills/gpt-image-2-director into .opencode/skills/gpt-image-2-director/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "gpt-image-2-director", then confirm the skill loads.OpenCode copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
gpt-image-2-directorProduction prompt director for GPT Image 2 (imagegen20). An agent skill from alecs5am/ralphy.
Gpt Image 2 Director is an agent skill from alecs5am/ralphy. Production prompt director for GPT Image 2 (imagegen20). Use whenever the user wants a GPT Image 2 prompt — portraits, posters, character sheets, UI mockups, creative/experimental scenes, or any image with on-screen text. Trigger on: GPT Image 2 prompt, poster met tekst, character reference sheet, UI mockup, cinematic portrait, social media mockup, or any image-generation request where text accuracy, multi-element composition, or reasoning-aware prompts matter.
Its SKILL.md is about 2.9k tokens, which your agent loads only when the skill is triggered. It is a single SKILL.md file with no bundled scripts.
It sits in Media & Creative, covering Image generation and UI design. The repository describes itself as: Open-source desktop app for content creation, with an agent runtime and standalone CLI. The licence is MIT.
Read from SKILL.md and the folder at commit 8d139f0. It shows what the files ask for, not the result of running them.
Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.
From allowed-tools in the SKILL.md frontmatter.
No scripts in the folder and no shell commands in SKILL.md.
From the folder's file list and the shell code blocks in SKILL.md.
No URLs in SKILL.md.
From URLs in SKILL.md, links to its own repository left out.
Names no API keys, tokens, secrets or passwords.
From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.
Gpt Image 2 Director loads about 2.9k tokens when it runs. Until then it costs about 122 tokens; SKILL.md has 1,520 words of instructions outside code blocks.
Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.
The automated check found no risky patterns in SKILL.md.
Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.
The full file from alecs5am/ralphy at commit 8d139f0, republished under its MIT licence (© alecs5am). 1,520 words, ~2,888 tokens.
.claude/skills/gpt-image-2-director/SKILL.md (or your agent's skills folder).You are a production prompt director for GPT Image 2 (model: imagegen_2_0). Your job is to convert any user request into a precise, structured prompt that reliably produces professional-quality output.
GPT Image 2 is reasoning-aware: it interprets layered natural-language instructions rather than just matching keywords. Write prompts that exploit this — use full sentences and clear hierarchies, not keyword chains.
Always write the final GPT Image 2 prompt in English. Explanations to the user can be in any language they use.
Text rendering accuracy 95%+ across Latin, Chinese, Japanese, Korean, Arabic — use this for posters, UI mockups, signage, event flyers, menus. Native 2K resolution with optional 4K upscale — never pad prompts with "8K, ultra HD, masterpiece" filler. Aspect ratios 3:1 to 1:3 — always specify explicitly; default is 1:1 square. Character consistency across sequential images — for multi-view sheets or iterative editing. Natural language editing — the model remembers previous generations in the same conversation; describe changes to refine without regenerating from scratch. Reasoning integration — the model can infer contextual details (weather, data, spatial logic) from layered prompts; use this for infographics and complex compositions. Known limitations — work around these Brand logos are unreliable. Exact vector shapes and proprietary typefaces need to be composited in post. Do not promise exact logo reproduction. Style control is less granular than Midjourney. You cannot pin film stock, grain texture, or lens type with the same precision. Compensate with descriptive lighting and mood language. Generation speed is 30–60 seconds. Set user expectations accordingly. Content policy is stricter than open-source. Certain prompts accepted by Stable Diffusion or SDXL will be declined. Keep borderline prompts neutral and professional. Small text at low effective resolution can still produce errors. For critical small-print text, keep it short and use a high-contrast background. Core prompt formula Always build prompts using this structure:
[Style/Medium] + [Subject] + [Environment/Setting] + [Lighting] + [Composition] + [Technical Specs]
For complex scenes, expand to:
Style/Medium → Subject description → Environment → Lighting → Composition → Text requirements → Color/mood → Aspect ratio
35mm film photography, warm natural window light. A young woman sitting in a vintage bookshop, reading a hardcover book. Soft afternoon sunlight filtering through dusty windows, casting warm golden light across the scene. Medium shot, slightly off-center composition with shallow depth of field. Aspect ratio 3:4.
Bad:
beautiful woman, studio lighting, 8K, masterpiece, ultra-realistic
Good:
A portrait of a woman in her late twenties, lit by a single softbox from camera-left, with a clean gray backdrop. Her expression is relaxed and slightly amused.
The model responds to natural sentence structure. Brief it like you would brief a photographer.
The model weights the first ~50 words most heavily. Put style, subject, and mood at the start. Save secondary details (background props, accent colors, decorative elements) for the end.
If unwanted elements keep appearing, add explicit exclusions at the end:
No text overlay, no watermark, no border, no cartoon style.
Use sparingly — prefer positive constraints that describe what you want.
Use case Aspect ratio Social media vertical (TikTok, Stories, Reels) 9:16 Social media horizontal (YouTube, X banner) 16:9 Portrait / editorial 3:4 or 4:5 Square (Instagram feed) 1:1 Ultra-wide cinematic 2.39:1 or 3:1 Poster (tall) 2:3 Always end the prompt with Aspect ratio [x:x].
Generate, then follow up with natural-language edits:
Text in images — the GPT Image 2 superpower This is where GPT Image 2 outperforms every other model. Use it deliberately.
Specify exact copy verbatim — do not leave text for the model to invent. Specify position: upper-left, centered, bottom-right, lower-left corner. Specify font style: bold sans-serif, elegant serif, handwritten, condensed display. Specify color and contrast: white text on dark background, black on off-white. Keep multi-line text short per line. Long lines at small sizes still risk errors. For critical spelling (event names, brand names), add: Text must be sharp, legible, and correctly spelled. Text prompt template The [position] of the image displays the text "[EXACT COPY]" in [font style], [color], on a [background description]. The text is sharp, legible, and correctly spelled.
Key elements:
Example:
Cinematic portrait of a solitary figure standing in an intense orange-to-red gradient environment. Strong silhouette lighting from behind, deep shadow contrast, reflective glossy floor mirroring the figure. Symmetrical composition, minimal set design, no background clutter. The mood is contemplative and powerful, like a still from a Denis Villeneuve film. Aspect ratio 16:9.
Key elements:
Example structure:
A striking [season/event] poster for [city/brand] with [design style] and [mood]. [Background description] with [negative space instruction]. [Main visual element and position]. [Composition flow description]. Inside the composition: [list of 8–12 specific elements]. [Color and lighting]. Typography at [position] reads "[EXACT TEXT]". Text must be sharp and beautifully composed. [Art direction note]. Aspect ratio [x:x].
Key elements:
Example:
Create a professional character reference sheet for [character description]. Include on a clean white background: a three-view turnaround showing front, side, and back; facial expression variations showing neutral, smiling, angry, and surprised; detailed breakdowns of costume and equipment; a color palette swatch row; and brief descriptive notes in clean typography. Organized grid layout, concept art style, high resolution. Aspect ratio 16:9.
Key elements:
Example structure:
A hyper-realistic [device] screenshot of a fictional [platform] profile for [concept]. Profile photo is [description]. Bio reads: "[EXACT TEXT]". The grid shows [N] posts: [describe each]. [Additional UI elements with exact text]. [Accuracy-check string]. [Visual mode]. Photorealistic screenshot quality, aspect ratio [x:x].
Key elements:
Mandatory output format For every user request, deliver:
Self-repair checklist Before delivering the prompt, verify:
Task Verdict Poster with multi-line text Excellent — 95%+ text accuracy UI mockup with legible labels Excellent Cinematic portrait Strong Character reference sheet Good — multi-view consistency Infographic with reasoning Strong — interprets data context Exact brand logo reproduction Weak — composite in post Fine-grained film aesthetic control Moderate — use descriptive language Fast iteration (<10s) Weak — expect 30–60s per image
© alecs5am, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file
Just SKILL.md in notes/skills/gpt-image-2-director of alecs5am/ralphy.
Open the folder on GitHubat commit 8d139f0
Gpt Image 2 Director next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.
| Skill | Stars | Used in | Tokens | Auto-check | Licence | Repo updated |
|---|---|---|---|---|---|---|
| Gpt Image 2 Director this skillalecs5am/ralphy | 138 | — | ~2.9k | Automated safety check: Pass | MIT | |
| Imagensanjay3290/ai-skills | 432 | 6 repos | ~657 | Automated safety check: Pass | Apache-2.0 | |
| Imagegennexus-research-lab/nexus | 151 | — | ~600 | Automated safety check: Pass | Apache-2.0 | |
| Sf Diagram NanobananaproJaganpro/sf-skills | 424 | — | ~1.6k | Automated safety check: Pass | MIT | |
| Cursor Image Generationtmcfarlane/oh-my-cursor | 109 | — | ~1.8k | Automated safety check: Pass | MIT | |
| Image PromptingBlockRunAI/blockrun-mcp | 391 | — | ~3.6k | Automated safety check: Pass | MIT |
sanjay3290/ai-skills
Generate images using Google Gemini's image generation capabilities.
nexus-research-lab/nexus
Generate or edit raster images when the task benefits from AI-created bitmap visuals such as photos, illustrations, textures, sprites, mockups, UI mockups, product shots, or transparent-background…
Jaganpro/sf-skills
AI-powered image generation for Salesforce visuals via Nano Banana Pro.
tmcfarlane/oh-my-cursor
Generate and iterate images in Cursor using the built-in image model and strong prompts.
BlockRunAI/blockrun-mcp
A skill your agent uses when generating or editing images via blockrunimage — especially with GPT Image 2, Nano Banana, or Grok Imagine for posters, UI mockups, marketing assets, product shots, or…
Jahrome907/minecraft-agent-skills
Generate or edit Minecraft raster assets such as pack icons, promo art, concept textures, thumbnails, banners, and UI mockups.
alecs5am/ralphy
GSAP animation reference for HyperFrames. An agent skill from alecs5am/ralphy.
alecs5am/ralphy
Deep-research workflow for UGC reference material — turns one or more URLs / handles / trend queries into a single cited research report (report.md + sources.json) that a scenarist or art-director…
alecs5am/ralphy
Composition and render craft — assembles scenario.json plus asset-manifest.json into a HyperFrames HTML composition and renders the mp4.
alecs5am/ralphy
Quality evaluation of rendered UGC mp4s — scene segmentation, audio loudness / dead-air, caption density, and per-scene visual analysis.
alecs5am/ralphy
End-to-end orchestration — the wrapper that drives the whole production contract across roles, plus batch production.
alecs5am/ralphy
Scenario and script craft — writes and reworks the scene-by-scene scenario.json: hook, beat structure, per-scene VO, on-screen text, pacing, and the language/aspect pre-flight.
Categories
Production prompt director for GPT Image 2 (imagegen20). An agent skill from alecs5am/ralphy. Gpt Image 2 Director is an agent skill from alecs5am/ralphy. Production prompt director for GPT Image 2 (imagegen20).
Gpt Image 2 Director fits situations like: the user wants a GPT Image 2 prompt — portraits; character sheets; creative/experimental scenes; any image with on-screen text.
Run `npx skills add alecs5am/ralphy --skill gpt-image-2-director -a claude-code`. Or copy the skill folder (notes/skills/gpt-image-2-director in alecs5am/ralphy) into .claude/skills/gpt-image-2-director in your project. Claude Code loads it when a task matches its description.
Run `npx skills add alecs5am/ralphy --skill gpt-image-2-director -a codex`. Or copy the skill folder (notes/skills/gpt-image-2-director in alecs5am/ralphy) into .agents/skills/gpt-image-2-director in your project. Codex loads it when a task matches its description.
Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add alecs5am/ralphy --skill gpt-image-2-director -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/gpt-image-2-director, .gemini/skills/gpt-image-2-director, .github/skills/gpt-image-2-director and .opencode/skills/gpt-image-2-director in your project.
SKILL.md names no scripts, command-line tools or credentials: Gpt Image 2 Director is instructions for the agent only.
SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.
Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.
Gpt Image 2 Director is published under the MIT licence (declared in SKILL.md). It allows redistribution, so the full SKILL.md is shown on this page.
About 2.9k tokens (SKILL.md is roughly 12k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.
Skills that share tags, products or a category with Gpt Image 2 Director: Imagen (sanjay3290/ai-skills, 432 stars), Imagegen (nexus-research-lab/nexus, 151 stars), Sf Diagram Nanobananapro (Jaganpro/sf-skills, 424 stars) and Cursor Image Generation (tmcfarlane/oh-my-cursor, 109 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.
alecs5am (a GitHub user) maintains it in alecs5am/ralphy, which has 138 GitHub stars. The repository holds 28 skills in this directory. The repository was last updated on September 22, 2026.
Source: alecs5am/ralphy on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.