Agent skill

Videoagent Image Studio

by pexoai in pexoai/pexo-skills

Tired of juggling 8 API keys?. An agent skill from pexoai/pexo-skills.

MITAuto-check passedMedia & Creative

Install Videoagent Image Studio

skills CLI
$ npx skills add pexoai/pexo-skills --skill videoagent-image-studio -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install pexoai/pexo-skills videoagent-image-studio --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/pexoai/pexo-skills.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/videoagent-image-studio .claude/skills/videoagent-image-studio && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
videoagent-image-studio
GitHub stars
802
Token cost
~1.7k tokens
SKILL.md length
518 words
Files
6
Skills in repo
18
Repo updated
First seen
Licence
MIT

At a glance

Tired of juggling 8 API keys?. An agent skill from pexoai/pexo-skills.

  • Works in 3 steps: Enhance the prompt → Run the script → Return the result
  • You want to generate any image without worrying about API keys
  • SKILL.md covers Quick Reference, How to Generate an Image, Midjourney Actions and Example Conversations, plus 2 more sections
  • Runs JavaScript scripts from its folder; calls node; reaches image-gen-proxy.vercel.app; needs IMAGE_STUDIO_TOKEN and FAL_KEY

What it does

Videoagent Image Studio is an agent skill from pexoai/pexo-skills. Tired of juggling 8 API keys? This skill gives you one-command access to Midjourney, Flux, Ideogram, and more, with zero setup. Use when you want to generate any image without worrying about API keys.

Its SKILL.md is about 1.7k tokens, which your agent loads only when the skill is triggered. The skill folder holds 6 other files (for example `CHANGELOG.md`, `CONTRIBUTING.md` and `package.json`).

It sits in Media & Creative, covering Image generation. It works with Midjourney. The repository describes itself as: A collection of open-source Agent Skills for content creation — images, audio, and video. The licence is MIT.

When your agent uses it

  • You want to generate any image without worrying about API keys
  • Tasks that involve Image generation

Example prompts

  • “/videoagent-image-studio”

Requirements

  • Node.js
  • A credential in IMAGE_STUDIO_TOKEN
  • A credential in FAL_KEY

Workflow steps

3 steps, taken from the step headings in SKILL.md.

  1. Enhance the prompt
  2. Run the script
  3. Return the result

What it can do on your machine

Read from SKILL.md and the folder at commit f724267. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Ships script files (JavaScript), which the agent can run.

    Shell commands in SKILL.md call:

    • node

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    Hosts in commands or code, which the agent is likely to contact:

    • image-gen-proxy.vercel.app

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names these keys or tokens, usually read from environment variables:

    • IMAGE_STUDIO_TOKEN
    • FAL_KEY
    • LEGNEXT_KEY

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Videoagent Image Studio loads about 1.7k tokens when it runs. Until then it costs about 56 tokens; SKILL.md has 518 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~56
When it runs · the whole SKILL.md, loaded when a task matches
~1.7k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from pexoai/pexo-skills at commit f724267, republished under its MIT licence (© pexoai). 518 words, ~1,686 tokens.

Download SKILL.mdSave it as .claude/skills/videoagent-image-studio/SKILL.md (or your agent's skills folder). This skill also uses 5 other files; get the full folder from GitHub.
name
videoagent-image-studio
description
Tired of juggling 8 API keys? This skill gives you one-command access to Midjourney, Flux, Ideogram, and more, with zero setup. Use when you want to generate any image without worrying about API keys.
version
2.0.0
author
wells
emoji
🎨
tags
video, image-generation, midjourney, flux, gemini, fal, ideogram, recraft
homepage
https://github.com/pexoai/image-studio-skill

🎨 VideoAgent Image Studio

Use when: User asks to generate, draw, create, or make any kind of image, photo, illustration, icon, logo, or artwork.

Generate images with 8 state-of-the-art AI models. This skill automatically picks the best model for the job and handles all the complexity — including Midjourney's async polling — so you can focus on the conversation.


Quick Reference

User IntentModelSpeed
Artistic, cinematic, painterlymidjourney~15s
Photorealistic, portrait, productflux-pro~8s
General purpose, balancedflux-dev~10s
Quick draft, fast iterationflux-schnell~2s
Image with text, logo, posterideogram~10s
Vector art, icon, flat designrecraft~8s
Anime, stylized illustrationsdxl~5s
Gemini-powered, consistent stylenano-banana~12s

How to Generate an Image

Step 1 — Enhance the prompt

Before calling the script, expand the user's prompt with style, lighting, and quality descriptors appropriate for the chosen model.

  • Midjourney: Add cinematic lighting, ultra detailed, --v 7, --style raw
  • Flux: Add masterpiece, highly detailed, sharp focus, professional photography
  • Ideogram: Be explicit about text content, font style, and layout
  • Recraft: Specify vector illustration, flat design, icon style
Step 2 — Run the script
bash
node {baseDir}/tools/generate.js \
  --model <model_id> \
  --prompt "<enhanced prompt>" \
  --aspect-ratio <ratio>

All parameters:

ParameterDefaultDescription
--modelflux-devModel ID from the table above
--prompt(required)The image generation prompt
--aspect-ratio1:11:1, 16:9, 9:16, 4:3, 3:4, 3:2, 21:9
--num-images1Number of images (1–4; Midjourney always returns 4)
--negative-prompt—Things to avoid (not supported by Midjourney)
--seed—Seed for reproducibility
Step 3 — Return the result

The script always waits and returns the final image URL(s). No polling required.

json
{
  "success": true,
  "model": "flux-pro",
  "imageUrl": "https://...",
  "images": ["https://..."]
}

Send the imageUrl to the user.


Midjourney Actions

After generating a 4-image grid with Midjourney, offer the user these options:

bash
# Upscale image #2 (subtle, preserves details)
node {baseDir}/tools/generate.js \
  --model midjourney \
  --action upscale \
  --index 2 \
  --job-id <job_id>

# Create a strong variation of image #3
node {baseDir}/tools/generate.js \
  --model midjourney \
  --action variation \
  --index 3 \
  --job-id <job_id> \
  --variation-type 1

# Regenerate with same prompt
node {baseDir}/tools/generate.js \
  --model midjourney \
  --action reroll \
  --job-id <job_id>

Upscale types: 0 = Subtle (default, best for photos), 1 = Creative (best for illustrations)

Variation types: 0 = Subtle (default), 1 = Strong (dramatic changes)


Show full SKILL.md (230 more words)Show less

Example Conversations

User: "Draw a snow leopard on a snowy mountain with cinematic lighting"

bash
# Choose midjourney for artistic quality
node {baseDir}/tools/generate.js \
  --model midjourney \
  --prompt "a majestic snow leopard on a snowy mountain peak, cinematic lighting, dramatic atmosphere, ultra detailed --ar 16:9 --v 7" \
  --aspect-ratio 16:9

🎨 Done! Which one to upscale? (U1-U4) Or create a variant? (V1-V4)


User: "Use Flux to generate a perfume product poster, white background"

bash
# Choose flux-pro for photorealistic product shots
node {baseDir}/tools/generate.js \
  --model flux-pro \
  --prompt "a luxury perfume bottle on a clean white background, professional product photography, soft shadows, 8k, highly detailed" \
  --aspect-ratio 3:4

User: "Show me a quick draft"

bash
# flux-schnell for instant previews
node {baseDir}/tools/generate.js \
  --model flux-schnell \
  --prompt "..." \
  --aspect-ratio 1:1

User: "Make me an App icon, flat style, blue theme"

bash
# recraft for vector/icon style
node {baseDir}/tools/generate.js \
  --model recraft \
  --prompt "a minimal flat design app icon, blue color scheme, simple geometric shapes, vector style, white background"

Setup

Zero API keys needed! All requests go through a hosted proxy that handles authentication server-side.

The skill works out of the box — just install and use.

Advanced: Custom proxy or token

If you want to use your own proxy or a persistent token, set these environment variables:

json
{
  "skills": {
    "entries": {
      "videoagent-image-studio": {
        "enabled": true,
        "env": {
          "IMAGE_STUDIO_PROXY_URL": "https://your-proxy.vercel.app",
          "IMAGE_STUDIO_TOKEN": "your_token_here"
        }
      }
    }
  }
}
VariableRequiredDescription
IMAGE_STUDIO_PROXY_URLNoCustom proxy base URL (default: https://image-gen-proxy.vercel.app)
IMAGE_STUDIO_TOKENNoPersistent token (auto-obtained if not set, 100 free uses per token)

To deploy your own proxy, see the videoagent-audio-studio proxy as a reference implementation. You'll need FAL_KEY and LEGNEXT_KEY as Vercel environment variables.


Changelog

v2.0.0
  • Simplified async: The script now blocks until Midjourney completes. No more --async / --poll flags needed in SKILL.md instructions.
  • Unified output format: All models return the same { success, imageUrl, images } shape.
  • Reference images for Nano Banana: Pass --reference-images "url1,url2" for character/style consistency across generations.
v1.3.0
  • Added non-blocking async mode for Midjourney (--async + --poll).
v1.2.0
  • Midjourney turbo mode enabled by default (~10-20s).
v1.1.0
  • Switched Midjourney provider from TTAPI to Legnext.ai for better stability.
v1.0.0
  • Initial release with Midjourney, Flux, SDXL, Nano Banana, Ideogram, Recraft.

© pexoai, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

SKILL.md and 5 other files in skills/videoagent-image-studio of pexoai/pexo-skills.

  • SKILL.md
  • .env.example
  • CHANGELOG.md
  • CONTRIBUTING.md
  • package.json
  • tools/generate.js

Open the folder on GitHubat commit f724267

Compare with similar skills

Videoagent Image Studio next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Videoagent Image Studio compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Videoagent Image Studio this skillpexoai/pexo-skills802—~1.7kAutomated safety check: PassMIT
Design Masterminhnv0807/ai-business-skills610—~4.6kAutomated safety check: PassMIT
AI Image Creative Directiondylanfeltus/skills179—~2.6kAutomated safety check: PassMIT
Character Design Sheet PromptRylaispirit/cinematic-video-prompt-skill147—~941Automated safety check: PassMIT
Image PromptKiyoraka/Project-AI-MemoryCore374—~2.4kAutomated safety check: PassNone
Image Postersanqiufong/slides-from-anything1321 repos~857Automated safety check: PassApache-2.0

Similar skills

  • Design Master

    minhnv0807/ai-business-skills

    Handles eight kinds of marketing visual requests, from logos and campaign key visuals to infographics and quote graphics, by generating images or writing paste-ready prompts.

    610 GitHub stars~4.6k tokensUpdated 27 days ago
    Media & CreativeAuto-check passed
  • AI Image Creative Direction

    dylanfeltus/skills

    Gives prompt templates, a model-selection guide and anti-generic rules for AI-generated visuals: hero images, feature illustrations, OG cards, icons and backgrounds.

    179 GitHub stars~2.6k tokensUpdated 21 days ago
    Media & CreativeAuto-check passed
  • Character Design Sheet Prompt

    Rylaispirit/cinematic-video-prompt-skill

    Writes one image prompt for a 16:9 character reference sheet, with three full-body views and a face close-up, so a character stays consistent across AI video clips.

    147 GitHub stars~941 tokensUpdated 12 days ago
    Media & CreativeAuto-check passed
  • Image Prompt

    Kiyoraka/Project-AI-MemoryCore

    Auto-triggers when user asks for a Midjourney or NijiJourney image prompt, when creating visual art prompts, or when user says 'midjourney prompt', 'niji prompt', 'create prompt', 'create a prompt'…

    374 GitHub stars~2.4k tokensUpdated 2 mo ago
    Media & CreativeAuto-check passed
  • Image Poster

    sanqiufong/slides-from-anything

    Single-image generation skill for posters, key art, and editorial illustrations.

    132 GitHub starsUsed in 1 repo~857 tokens
    Media & CreativeAuto-check passed
  • Image Prompt

    social-media-skills/skills

    Use as the foundation and router for any image a post needs — turn a social need into a clear image brief, describe it with model-agnostic prompt craft, and route to the right tool.

    128 GitHub stars~1.6k tokensUpdated 7 days ago
    Media & CreativeAuto-check passed

More from pexoai/pexo-skills

All 18 skills in this repo
  • AI Video Generation

    pexoai/pexo-skills

    Generate AI video from any input — text, image, or script — with Pexo.

    802 GitHub stars~1.3k tokensUpdated 1 mo ago
    Auto-check passed
  • Explainer Video

    pexoai/pexo-skills

    Create an explainer video with narration using Pexo. An agent skill from pexoai/pexo-skills.

    802 GitHub stars~1.3k tokensUpdated 1 mo ago
    Auto-check passed
  • Founder Video

    pexoai/pexo-skills

    Make a founder video with Pexo — built for solo founders and small teams.

    802 GitHub stars~1.4k tokensUpdated 1 mo ago
    Auto-check passed
  • Image To Video

    pexoai/pexo-skills

    Animate a still image into a finished, moving video with Pexo.

    802 GitHub stars~1.3k tokensUpdated 1 mo ago
    Auto-check passed
  • Launch Video

    pexoai/pexo-skills

    Make a launch video for your startup or product with Pexo. An agent skill from pexoai/pexo-skills.

    802 GitHub stars~1.4k tokensUpdated 1 mo ago
    Auto-check passed
  • Make A Video

    pexoai/pexo-skills

    Make a complete video from a simple idea with Pexo. An agent skill from pexoai/pexo-skills.

    802 GitHub stars~1.3k tokensUpdated 1 mo ago
    Auto-check passed

Works with

Questions about Videoagent Image Studio

What does Videoagent Image Studio do?

Tired of juggling 8 API keys?. An agent skill from pexoai/pexo-skills. Videoagent Image Studio is an agent skill from pexoai/pexo-skills. Tired of juggling 8 API keys?

When should I use Videoagent Image Studio?

Videoagent Image Studio fits situations like: you want to generate any image without worrying about API keys; tasks that involve Image generation.

How do I install Videoagent Image Studio in Claude Code?

Run `npx skills add pexoai/pexo-skills --skill videoagent-image-studio -a claude-code`. Or copy the skill folder (skills/videoagent-image-studio in pexoai/pexo-skills) into .claude/skills/videoagent-image-studio in your project. Claude Code loads it when a task matches its description.

How do I install Videoagent Image Studio in Codex?

Run `npx skills add pexoai/pexo-skills --skill videoagent-image-studio -a codex`. Or copy the skill folder (skills/videoagent-image-studio in pexoai/pexo-skills) into .agents/skills/videoagent-image-studio in your project. Codex loads it when a task matches its description.

Can I use Videoagent Image Studio in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add pexoai/pexo-skills --skill videoagent-image-studio -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/videoagent-image-studio, .gemini/skills/videoagent-image-studio, .github/skills/videoagent-image-studio and .opencode/skills/videoagent-image-studio in your project.

What does Videoagent Image Studio need to run?

Going by SKILL.md and its folder, Videoagent Image Studio needs JavaScript for the scripts in its folder, the command-line tools its instructions call (node) and credentials named IMAGE_STUDIO_TOKEN, FAL_KEY and LEGNEXT_KEY. Our summary lists: Node.js; A credential in IMAGE_STUDIO_TOKEN; A credential in FAL_KEY.

Does Videoagent Image Studio access the network?

SKILL.md names 1 domain. In commands or code: image-gen-proxy.vercel.app; the agent is likely to contact it when it follows the instructions. This is read from the text; nothing was executed.

Is Videoagent Image Studio safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Videoagent Image Studio use?

Videoagent Image Studio is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Videoagent Image Studio use?

About 1.7k tokens (SKILL.md is roughly 6.7k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Videoagent Image Studio?

Skills that share tags, products or a category with Videoagent Image Studio: Design Master (minhnv0807/ai-business-skills, 610 stars), AI Image Creative Direction (dylanfeltus/skills, 179 stars), Character Design Sheet Prompt (Rylaispirit/cinematic-video-prompt-skill, 147 stars) and Image Prompt (Kiyoraka/Project-AI-MemoryCore, 374 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Videoagent Image Studio?

pexoai (a GitHub user) maintains it in pexoai/pexo-skills, which has 802 GitHub stars. The repository holds 18 skills in this directory. The repository was last updated on August 20, 2026.

Source: pexoai/pexo-skills on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.