Agent skill

AI Image Creative Direction

by dylanfeltus in dylanfeltus/skills

Gives prompt templates, a model-selection guide and anti-generic rules for AI-generated visuals: hero images, feature illustrations, OG cards, icons and backgrounds.

MITAuto-check passedMedia & Creative

Install AI Image Creative Direction

skills CLI
$ npx skills add dylanfeltus/skills --skill creative-direction -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install dylanfeltus/skills creative-direction --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/dylanfeltus/skills.git skills-src && mkdir -p .claude/skills && cp -r skills-src/creative-direction .claude/skills/creative-direction && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
creative-direction
GitHub stars
179
Token cost
~2.6k tokens
SKILL.md length
903 words
Files
2
Skills in repo
10
Repo updated
First seen
Licence
MIT

At a glance

Gives prompt templates, a model-selection guide and anti-generic rules for AI-generated visuals: hero images, feature illustrations, OG cards, icons and backgrounds.

  • Works in 4 steps: Generic: "A workspace" ❌ → Specific: "A designer's workspace with a… → Atmospheric: "A designer's workspace at… → …
  • Choosing an image model for a landing page or marketing asset
  • SKILL.md covers When to Use, Core Philosophy, Model Selection Guide and Prompt Templates by Asset Type, plus 4 more sections
  • Instructions only: no scripts, shell commands, URLs or credentials in SKILL.md

What it does

The skill's core philosophy is that generic images are forgettable and every image has a job: a hero image carries emotion and aspiration, a feature image carries clarity, a testimonial carries trust. It argues that a consistent set of good-enough images beats one perfect image surrounded by mismatched ones.

A model selection guide compares GPT-4o and DALL-E 3 for text and precise composition, Gemini Imagen for photorealism and product shots, Midjourney for mood and cinematic quality without an API, Flux via Replicate for photorealistic people, and stock photo sites when AI looks too artificial, with a short decision tree for picking between them.

Prompt templates are broken out by asset type, starting with the hero image: a template of art style, a specific aspirational subject action, mood-setting environment details, lighting, a color palette constraint and a composition note, illustrated with a SaaS product example.

When your agent uses it

  • Choosing an image model for a landing page or marketing asset
  • Writing a prompt for a hero image, feature illustration or OG card
  • Fixing AI-generated images that look generic or stock-photo-like

Example prompts

  • “Write a hero image prompt for our SaaS dashboard product, aiming for a cinematic morning office scene.”
  • “Which model should I use to generate a realistic OG card with text on it?”
  • “These feature icons look too generic. Give me a more specific prompt.”

Workflow steps

4 steps, taken from the first numbered list in SKILL.md.

  1. Generic: "A workspace" ❌
  2. Specific: "A designer's workspace with a drawing tablet" ⬆️
  3. Atmospheric: "A designer's workspace at golden hour, warm light on a Wacom tablet" ⬆️
  4. Story: "A designer's workspace at golden hour, a half-finished illustration on the tablet, coffee cup with a lipstick mark, headphones…

What it can do on your machine

Read from SKILL.md and the folder at commit b97a48f. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    No scripts in the folder and no shell commands in SKILL.md.

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

AI Image Creative Direction loads about 2.6k tokens when it runs. Until then it costs about 72 tokens; SKILL.md has 903 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~72
When it runs · the whole SKILL.md, loaded when a task matches
~2.6k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from dylanfeltus/skills at commit b97a48f, republished under its MIT licence (© dylanfeltus). 903 words, ~2,648 tokens.

Download SKILL.mdSave it as .claude/skills/creative-direction/SKILL.md (or your agent's skills folder). This skill also uses 1 other file; get the full folder from GitHub.
name
creative-direction
description
Image prompt templates, model selection guidance, and anti-generic patterns for generating visual assets. Use when the user needs AI-generated images for landing pages, marketing, or products. Covers hero images, feature illustrations, OG cards, icons, and backgrounds.

Creative Direction

Image prompt templates, model selection guidance, and anti-generic patterns for generating visual assets. Covers hero images, feature illustrations, testimonial photos, OG images, and more.

When to Use

  • User needs images for a landing page, app, or marketing site
  • User asks for "hero image", "feature illustration", or "OG image"
  • User wants AI-generated visuals that don't look like stock photos
  • User is choosing between image generation models
  • User's current AI images look generic and need direction

Core Philosophy

  • Generic is the enemy. "A person using a laptop" produces forgettable images. Specificity creates memorability.
  • Every image has a job. Hero = emotion + aspiration. Feature = clarity. Testimonial = trust. Know the job before prompting.
  • Model matters. Different models excel at different things. Pick the right tool.
  • Consistency beats quality. A cohesive set of 7/10 images beats one 10/10 with six mismatches.

Model Selection Guide

When to Use Each Model
ModelBest ForWeaknessesCost
GPT-4o / DALL-E 3Text in images, diagrams, infographics, precise compositionsCan feel "illustrated", less photorealisticAPI credits
Gemini ImagenPhotorealism, natural scenes, product shotsLess control over composition, text rendering variesAPI credits
MidjourneyArtistic quality, mood, cinematic shots, brand imageryNo API (Discord-only), inconsistent with specific detailsSubscription
Flux (via Replicate)Photorealism, faces, flexible stylesRequires Replicate accountPer-image
Unsplash / PexelsReal photography, when AI looks too AILimited to what exists, generic posesFree
Decision Tree
Need text in the image? → DALL-E 3 / GPT-4o
Need photorealistic people? → Flux or Gemini
Need artistic/cinematic mood? → Midjourney
Need a specific real-world scene? → Unsplash/Pexels
Need consistency across many images? → Pick ONE model, same style prompt prefix
Need diagrams/UI mockups? → DALL-E 3 / GPT-4o

Prompt Templates by Asset Type

Hero Image

Job: Create an emotional first impression. Communicate the product's vibe in 2 seconds.

Template:

[Art style], [subject doing something specific and aspirational],
[environment with mood-setting details], [lighting description],
[color palette constraint], [composition note]

Example — SaaS Product:

Cinematic photograph, a product designer reviewing a clean dashboard
on a large monitor in a sunlit corner office, morning golden hour
light casting long shadows, muted blue and warm cream color palette,
shot from over the shoulder with shallow depth of field, 35mm lens feel

Anti-generic patterns:

  • ❌ "A person using a computer" → ✅ "A designer reviewing analytics on a ultrawide monitor, sticky notes scattered on the desk"
  • ❌ "Happy team working" → ✅ "Three engineers around a whiteboard, one mid-laugh, marker in hand, late afternoon light"
  • ❌ "Technology abstract" → ✅ "Closeup of hands arranging physical cards on a table, each card showing a tiny wireframe sketch"
Feature/Product Illustration

Job: Explain what a feature does. Clarity > beauty.

Template:

[Clean/minimal style], [the feature's core concept as a visual metaphor],
[simple background], [brand colors], [no text unless needed]

Example — Analytics Feature:

Minimal 3D illustration, a translucent glass cube containing floating
data points that form a rising trend line, soft gradient background
from white to light blue, subtle shadows, isometric perspective

Tips:

  • Use visual metaphors, not literal screenshots
  • Keep backgrounds simple — the illustration should work on any page section
  • Maintain consistent style across all feature illustrations (same rendering style, same perspective, same color treatment)
Testimonial / Social Proof

Job: Build trust. Make real people feel real.

Approach: Use real photos when possible (with permission). If generating:

Professional headshot, [specific person description with age/style details],
[neutral or office background], [natural lighting], [warm but professional mood],
shot at eye level, slight smile, [avoid uncanny valley — add imperfections]

Tips:

  • Diversity matters — vary age, ethnicity, style
  • Avoid the "corporate headshot" look — slightly candid feels more trustworthy
  • ⚠️ AI-generated faces for testimonials is ethically questionable. Prefer real photos. If generating, be transparent about it.
Open Graph (OG) / Social Card

Job: Get clicks in a feed. Must work at small sizes.

Template:

[Bold, high contrast], [simple central element],
[large readable text area (left or center)],
[brand color background or gradient], [16:9 aspect ratio],
minimal detail — this will be viewed at 600×315px

Key constraints:

  • Must be readable at thumbnail size
  • Text should be generated separately and composited (AI text rendering is unreliable)
  • High contrast between background and text area
  • Simple shapes > complex scenes
Icon / Logo Concept

Job: Convey brand identity in a tiny space.

Minimal vector icon, [object/symbol], [single or two-color],
clean lines, works at 32px, [style: geometric/rounded/sharp],
white background, no gradients, no shadows

Tips:

  • Generate concepts, then recreate in Figma/SVG for production
  • AI-generated logos are starting points, not finals
  • Test at small sizes — if it doesn't read at 32px, simplify
Background / Texture

Job: Add depth without competing with content.

Abstract [texture type], [color palette], subtle variation,
tileable/seamless, low contrast, [usage: dark background with light text / light background with dark text]

Texture types: gradient mesh, noise grain, geometric pattern, organic shapes, topographic lines, dot grid


Style Consistency Framework

When generating multiple images for a project, create a style prefix and prepend it to every prompt:

STYLE PREFIX: "Minimal 3D illustration, soft matte materials, isometric perspective,
pastel color palette with [brand blue] accents, subtle ambient occlusion shadows,
white background —"

Then each prompt becomes:

[STYLE PREFIX] a shield icon representing security features
[STYLE PREFIX] a speedometer showing performance optimization
[STYLE PREFIX] a connected graph showing team collaboration
Show full SKILL.md (362 more words)Show less
Consistency Checklist
  • Same art style (3D, flat, photographic, etc.)
  • Same color palette (or subset of it)
  • Same perspective (isometric, front-facing, etc.)
  • Same lighting direction
  • Same level of detail/complexity
  • Same background treatment

Anti-Generic Playbook

The Specificity Ladder

Each level adds memorability:

  1. Generic: "A workspace" ❌
  2. Specific: "A designer's workspace with a drawing tablet" ⬆️
  3. Atmospheric: "A designer's workspace at golden hour, warm light on a Wacom tablet" ⬆️
  4. Story: "A designer's workspace at golden hour, a half-finished illustration on the tablet, coffee cup with a lipstick mark, headphones draped over the monitor" ✅

Always aim for level 3–4.

Overused AI Image Tropes to Avoid
❌ Cliché✅ Alternative
Glowing orbs / particlesPhysical textures, natural materials
Floating holographic UIReal devices, paper prototypes
Purple/blue gradient everythingEarth tones, brand-specific palettes
Isometric city blocksFocused single-object compositions
Perfect symmetryIntentional asymmetry, rule of thirds
Hyper-saturated colorsMuted, desaturated palette with one accent
"AI art style" shininessMatte materials, film grain, imperfection
Adding Realism to AI Images

Include in prompts:

  • Film grain: "slight film grain, shot on 35mm"
  • Imperfection: "slightly worn edges", "coffee stain on the desk"
  • Natural lighting: "overcast diffused light" instead of "bright studio lighting"
  • Depth of field: "shallow depth of field, f/1.8" for focus
  • Texture: "matte finish", "linen texture", "concrete surface"

Output Format

When providing creative direction, output:

### Creative Direction: [Asset Type]

**Purpose:** What this image needs to communicate
**Model recommendation:** [Model] — [Why]
**Style:** [Art direction notes]

**Prompt:**
[Full prompt ready to paste]

**Variations to try:**
1. [Alternative angle/mood]
2. [Alternative style]

**Post-processing notes:**
- [Any needed adjustments — cropping, overlay, text addition]

Examples

Example 1: "I need a hero image for a project management SaaS"

Purpose: Communicate clarity and control over complex projects Model: Midjourney (cinematic quality) or DALL-E 3 (if text needed)

Prompt:

Cinematic overhead photograph of a large wooden desk with neatly organized
project cards, color-coded sticky notes in a kanban layout, a MacBook
showing a clean dashboard, a ceramic mug, natural window light from the left,
shallow depth of field focusing on the cards, muted warm palette with
one accent color (brand blue), 35mm film aesthetic with slight grain
Example 2: "Generate feature illustrations for our 4 main features"

Style prefix:

Minimal 3D illustration, soft matte clay-like materials, front-facing
perspective, brand indigo (#6366F1) as accent, light gray (#F9FAFB)
background, gentle directional shadow to the bottom-right —

Prompts:

  1. [prefix] a magnifying glass hovering over a organized grid of documents
  2. [prefix] two puzzle pieces connecting, with a small spark at the join
  3. [prefix] a clock face with segments in different colors showing time allocation
  4. [prefix] a shield with a small checkmark, slightly tilted
Example 3: "Our AI images look too generic, help"

Audit current images against the anti-generic playbook. Common fixes:

  1. Add specificity (level 3–4 on the ladder)
  2. Replace cliché tropes (glowing orbs → physical textures)
  3. Add film grain / imperfection to prompts
  4. Constrain the color palette
  5. Use consistent style prefix across all images

© dylanfeltus, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

SKILL.md and 1 other file in creative-direction of dylanfeltus/skills.

  • SKILL.md
  • README.md

Open the folder on GitHubat commit b97a48f

Compare with similar skills

AI Image Creative Direction next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

AI Image Creative Direction compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
AI Image Creative Direction this skilldylanfeltus/skills179—~2.6kAutomated safety check: PassMIT
Design Masterminhnv0807/ai-business-skills608—~4.6kAutomated safety check: PassMIT
Image Promptsocial-media-skills/skills116—~1.6kAutomated safety check: PassMIT
Ky Design To HTMLKyrieCheungYep/ky-design-to-html-skill163—~2.8kAutomated safety check: PassNone
Character Design Sheet PromptRylaispirit/cinematic-video-prompt-skill146—~941Automated safety check: PassMIT
Image PromptKiyoraka/Project-AI-MemoryCore374—~2.4kAutomated safety check: PassNone

Similar skills

  • Design Master

    minhnv0807/ai-business-skills

    Handles eight kinds of marketing visual requests, from logos and campaign key visuals to infographics and quote graphics, by generating images or writing paste-ready prompts.

    608 GitHub stars~4.6k tokensUpdated 25 days ago
    Media & CreativeAuto-check passed
  • Image Prompt

    social-media-skills/skills

    Use as the foundation and router for any image a post needs — turn a social need into a clear image brief, describe it with model-agnostic prompt craft, and route to the right tool.

    116 GitHub stars~1.6k tokensUpdated 5 days ago
    Media & CreativeAuto-check passed
  • Ky Design To HTML

    KyrieCheungYep/ky-design-to-html-skill

    Convert visible UI references such as screenshots, exported mockups, Figma frame images, SaaS empty states, dashboards, and landing page screenshots into high-quality HTML/CSS by decomposing layout…

    163 GitHub stars~2.8k tokensUpdated 4 mo ago
    Frontend & DesignAuto-check passed
  • Character Design Sheet Prompt

    Rylaispirit/cinematic-video-prompt-skill

    Writes one image prompt for a 16:9 character reference sheet, with three full-body views and a face close-up, so a character stays consistent across AI video clips.

    146 GitHub stars~941 tokensUpdated 9 days ago
    Media & CreativeAuto-check passed
  • Image Prompt

    Kiyoraka/Project-AI-MemoryCore

    Auto-triggers when user asks for a Midjourney or NijiJourney image prompt, when creating visual art prompts, or when user says 'midjourney prompt', 'niji prompt', 'create prompt', 'create a prompt'…

    374 GitHub stars~2.4k tokensUpdated 2 mo ago
    Media & CreativeAuto-check passed
  • Omd Media

    kwakseongjae/oh-my-design

    Brand-consistent asset sets from DESIGN.md — hero, section illustrations, icon sets, OG image — generated through whichever channel the user actually has (grok build imagegen, Codex $imagegen…

    531 GitHub stars~1.1k tokensUpdated 6 days ago
    Frontend & DesignAuto-check passed

More from dylanfeltus/skills

All 10 skills in this repo
  • App Store Intelligence

    dylanfeltus/skills

    Looks up app details, ratings and reviews and searches the iOS and Mac App Stores through Apple's free iTunes APIs, with a web-search fallback for Google Play.

    179 GitHub stars~2k tokensUpdated 19 days ago
    Auto-check passed
  • Design Token Generator

    dylanfeltus/skills

    Generates type scales, color palettes, spacing systems, WCAG contrast checks and dark mode palettes from formulas, as CSS custom properties, a Tailwind config or JSON tokens.

    179 GitHub stars~2.9k tokensUpdated 19 days ago
    Auto-check passed
  • Hacker News Search

    dylanfeltus/skills

    Searches and monitors Hacker News stories, comments and users through the free Algolia HN Search API, with tag, points and date filters.

    179 GitHub stars~1.7k tokensUpdated 19 days ago
    Auto-check passed
  • Framer Motion (Motion) patterns for React: spring presets, staggers, layout animations, micro-interactions, scroll effects and page transitions, with performance rules.

    179 GitHub stars~3k tokensUpdated 19 days ago
    Auto-check passed
  • Privacy.com Virtual Cards

    dylanfeltus/skills

    Creates, updates, and monitors Privacy.com virtual payment cards with spending limits for an agent making controlled purchases.

    179 GitHub stars~2.2k tokensUpdated 19 days ago
    Auto-check passed
  • Looks up Product Hunt launches, products and makers through its GraphQL API: daily top posts, topic browsing and per-post details, using a free developer token.

    179 GitHub stars~1.7k tokensUpdated 19 days ago
    Auto-check passed

Works with

Questions about AI Image Creative Direction

What does AI Image Creative Direction do?

Gives prompt templates, a model-selection guide and anti-generic rules for AI-generated visuals: hero images, feature illustrations, OG cards, icons and backgrounds. The skill's core philosophy is that generic images are forgettable and every image has a job: a hero image carries emotion and aspiration, a feature image carries clarity, a testimonial carries trust. It argues that a consistent set of good-enough images beats one perfect image surrounded by mismatched ones.

When should I use AI Image Creative Direction?

AI Image Creative Direction fits situations like: choosing an image model for a landing page or marketing asset; writing a prompt for a hero image, feature illustration or OG card; fixing AI-generated images that look generic or stock-photo-like.

How do I install AI Image Creative Direction in Claude Code?

Run `npx skills add dylanfeltus/skills --skill creative-direction -a claude-code`. Or copy the skill folder (creative-direction in dylanfeltus/skills) into .claude/skills/creative-direction in your project. Claude Code loads it when a task matches its description.

How do I install AI Image Creative Direction in Codex?

Run `npx skills add dylanfeltus/skills --skill creative-direction -a codex`. Or copy the skill folder (creative-direction in dylanfeltus/skills) into .agents/skills/creative-direction in your project. Codex loads it when a task matches its description.

Can I use AI Image Creative Direction in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add dylanfeltus/skills --skill creative-direction -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/creative-direction, .gemini/skills/creative-direction, .github/skills/creative-direction and .opencode/skills/creative-direction in your project.

What does AI Image Creative Direction need to run?

SKILL.md names no scripts, command-line tools or credentials: AI Image Creative Direction is instructions for the agent only.

Does AI Image Creative Direction access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is AI Image Creative Direction safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does AI Image Creative Direction use?

AI Image Creative Direction is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does AI Image Creative Direction use?

About 2.6k tokens (SKILL.md is roughly 11k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to AI Image Creative Direction?

Skills that share tags, products or a category with AI Image Creative Direction: Design Master (minhnv0807/ai-business-skills, 608 stars), Image Prompt (social-media-skills/skills, 116 stars), Ky Design To HTML (KyrieCheungYep/ky-design-to-html-skill, 163 stars) and Character Design Sheet Prompt (Rylaispirit/cinematic-video-prompt-skill, 146 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains AI Image Creative Direction?

dylanfeltus (a GitHub user) maintains it in dylanfeltus/skills, which has 179 GitHub stars. The repository holds 10 skills in this directory. The repository was last updated on September 17, 2026.

Source: dylanfeltus/skills on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.