Agent skill

Videoagent Video Studio

by pexoai in pexoai/pexo-skills

Generate short AI videos from text or images — text-to-video, image-to-video, and reference-based generation — with zero API key setup.

MITAuto-check passedMedia & Creative

Install Videoagent Video Studio

skills CLI
$ npx skills add pexoai/pexo-skills --skill videoagent-video-studio -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install pexoai/pexo-skills videoagent-video-studio --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/pexoai/pexo-skills.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/videoagent-video-studio .claude/skills/videoagent-video-studio && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
videoagent-video-studio
GitHub stars
804
Token cost
~1.6k tokens
SKILL.md length
454 words
Files
25 (incl. scripts, references)
Skills in repo
18
Repo updated
First seen
Licence
MIT

At a glance

Generate short AI videos from text or images — text-to-video, image-to-video, and reference-based generation — with zero API key setup.

  • Works in 3 steps: Choose mode and enhance the prompt → Run the script → Return the result
  • The user wants to create a video clip
  • SKILL.md covers Quick Reference, How to Generate a Video, Example Conversations and Setup, plus 1 more section
  • Runs JavaScript scripts from its folder; calls node; needs VIDEO_STUDIO_TOKEN

What it does

Videoagent Video Studio is an agent skill from pexoai/pexo-skills. Generate short AI videos from text or images — text-to-video, image-to-video, and reference-based generation — with zero API key setup. Use when the user wants to create a video clip, animate an image, or generate video from a description.

Its SKILL.md is about 1.6k tokens, which your agent loads only when the skill is triggered. The skill folder holds 27 other files, including scripts and reference files (for example `README.md`, `package-lock.json` and `package.json`).

It sits in Media & Creative, covering AI video generation. It works with Seedance and MiniMax. The repository describes itself as: A collection of open-source Agent Skills for content creation — images, audio, and video. The licence is MIT.

When your agent uses it

  • The user wants to create a video clip
  • Animate an image
  • Generate video from a description

Example prompts

  • “/videoagent-video-studio”

Requirements

  • Node.js
  • A credential in VIDEO_STUDIO_TOKEN

Workflow steps

3 steps, taken from the step headings in SKILL.md.

  1. Choose mode and enhance the prompt
  2. Run the script
  3. Return the result

What it can do on your machine

Read from SKILL.md and the folder at commit f724267. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Ships 1 file in scripts/ (JavaScript, from the files we listed), which the agent can run.

    Shell commands in SKILL.md call:

    • node

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names these keys or tokens, usually read from environment variables:

    • VIDEO_STUDIO_TOKEN

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Videoagent Video Studio loads about 1.6k tokens when it runs, and up to ~4.3k if it reads all its reference files. Until then it costs about 66 tokens; SKILL.md has 454 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~66
When it runs · the whole SKILL.md, loaded when a task matches
~1.6k
With references · SKILL.md plus every file in references/, read only if the agent opens them
~4.3k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); the scripts in this folder are not scanned.

SKILL.md

The full file from pexoai/pexo-skills at commit f724267, republished under its MIT licence (© pexoai). 454 words, ~1,611 tokens.

Download SKILL.mdSave it as .claude/skills/videoagent-video-studio/SKILL.md (or your agent's skills folder). This skill also uses 24 other files; get the full folder from GitHub.
name
videoagent-video-studio
description
Generate short AI videos from text or images — text-to-video, image-to-video, and reference-based generation — with zero API key setup. Use when the user wants to create a video clip, animate an image, or generate video from a description.
version
2.1.0
author
pexoai
emoji
🎬
tags
video, video-generation, text-to-video, image-to-video, veo, grok, kling, seedance, minimax, hunyuan, pixverse

🎬 VideoAgent Video Studio

Use when: User asks to generate a video, create a video from text, animate an image, make a short clip, or produce AI video.

Generate short AI videos with 7 backends. This skill picks the right mode (text-to-video or image-to-video), enhances the prompt for best results, and returns the video URL.


Quick Reference

User IntentModeTypical Duration
"Make a video of..." (no image)text-to-video4–10 s
"Animate this image" / "Make this move"image-to-video4–6 s
"Turn this into a video with..."image-to-video4–6 s
Cinematic, story, adPrefer text-to-video with detailed prompt5–10 s
Generation Modes
ModeDescriptionModels
text-to-videoText prompt only → videominimax, kling, veo, hunyuan, grok, seedance
image-to-videoSingle image + prompt → animated clipminimax, kling, veo, pixverse, grok, seedance
reference-basedReference images/video → consistent outputminimax, kling, veo, hunyuan, grok, seedance
Models (use --model <id>)
Model IDT2VI2VReferenceNotes
minimax✅✅✅Subject reference image, character consistency
kling✅✅✅Multi-element / character / keyframe (O3)
veo✅✅✅Google Veo 3.1, multiple reference images
hunyuan✅—✅Video-to-video style transfer
pixverse—✅—Stylized image-to-video
grok✅✅✅Video editing via reference video
seedance✅✅✅Seedance 1.5 Pro, synchronized audio, 4–12 s

Full model details and endpoint reference: references/models.md.


How to Generate a Video

Step 1 — Choose mode and enhance the prompt
  • Text-to-video: Expand with subject, action, camera movement, lighting, and style. Be specific about motion (e.g. "camera slowly zooms in", "character walks left to right").
  • Image-to-video: Describe the motion to apply to the image (e.g. "gentle breeze in the hair", "camera pans across the scene"). See references/prompt_guide.md for patterns.
Show full SKILL.md (200 more words)Show less
Step 2 — Run the script

Text-to-video:

bash
node {baseDir}/tools/generate.js \
  --mode text-to-video \
  --prompt "<enhanced prompt>" \
  --duration <seconds> \
  --aspect-ratio <ratio>

Image-to-video:

bash
node {baseDir}/tools/generate.js \
  --mode image-to-video \
  --prompt "<motion description>" \
  --image-url "<public image URL>" \
  --duration <seconds> \
  --aspect-ratio <ratio>

Parameters:

ParameterDefaultDescription
--modetext-to-videotext-to-video or image-to-video
--prompt(required)Scene or motion description
--image-url—Required for image-to-video; public image URL
--duration5Length in seconds (typically 4–10)
--aspect-ratio16:916:9, 9:16, 1:1, 4:3, 3:4
--modelautoModel ID (e.g. kling, veo, grok, seedance); auto = proxy picks

Other commands:

CommandDescription
node tools/generate.js --list-modelsList available models from the proxy
node tools/generate.js --status --job-id <id>Check async job status
Step 3 — Return the result

The script returns JSON:

json
{
  "success": true,
  "mode": "text-to-video",
  "videoUrl": "https://...",
  "duration": 5,
  "aspectRatio": "16:9"
}

Send videoUrl to the user.


Example Conversations

User: "Generate a short video of a cat walking in the rain, cinematic."

bash
node {baseDir}/tools/generate.js \
  --mode text-to-video \
  --prompt "A cat walking through rain, wet streets, neon reflections, cinematic lighting, slow motion, 4K" \
  --duration 5 \
  --aspect-ratio 16:9

User: "Animate this photo" (user uploads a landscape)

bash
node {baseDir}/tools/generate.js \
  --mode image-to-video \
  --prompt "Gentle clouds moving across the sky, subtle grass movement, cinematic atmosphere" \
  --image-url "https://..." \
  --duration 5 \
  --aspect-ratio 16:9

User: "Make a 10-second vertical video of a coffee pour, slow motion."

bash
node {baseDir}/tools/generate.js \
  --mode text-to-video \
  --prompt "Close-up of coffee pouring into a white cup, slow motion, steam rising, soft lighting, product shot" \
  --duration 10 \
  --aspect-ratio 9:16

User: "Use Google Veo for a cinematic shot."

bash
node {baseDir}/tools/generate.js \
  --mode text-to-video \
  --model veo \
  --prompt "A dragon flying through cloudy skies, cinematic lighting, 8s" \
  --duration 8 \
  --aspect-ratio 16:9

User: "Animate this portrait."

bash
node {baseDir}/tools/generate.js \
  --mode image-to-video \
  --model grok \
  --prompt "Gentle smile, subtle head turn" \
  --image-url "https://..." \
  --duration 5

Setup

Zero API keys by default. Requests go through a hosted proxy. Set these for a custom proxy or token:

VariableRequiredDescription
VIDEO_STUDIO_PROXY_URLNoProxy base URL
VIDEO_STUDIO_TOKENNoAuth token if the proxy requires it

Knowledge Base

© pexoai, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

SKILL.md and 24 other files (scripts, references) in skills/videoagent-video-studio of pexoai/pexo-skills.

  • SKILL.md
  • README.md
  • package-lock.json
  • package.json
  • proxy/README.md
  • proxy/api/generate.js
  • proxy/api/stats.js
  • proxy/api/status.js
  • proxy/api/token.js
  • proxy/models.js
  • proxy/package-lock.json
  • proxy/package.json
  • proxy/usage-store.js
  • proxy/vercel.json
  • references/calling_guide.md
  • references/models.md
  • references/prompt_guide.md
  • scripts
  • … and 7 more

Open the folder on GitHubat commit f724267

Compare with similar skills

Videoagent Video Studio next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Videoagent Video Studio compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Videoagent Video Studio this skillpexoai/pexo-skills804—~1.6kAutomated safety check: PassMIT
HiggsfieldOSideMedia/higgsfield-ai-prompt-skill713—~9.1kAutomated safety check: PassMIT
VideoNexus-JPF/note-companion8703 repos~3.6kAutomated safety check: PassMIT
Video PromptingSquare-Zero-Labs/video-prompting-skill182—~1.9kAutomated safety check: PassApache-2.0
AI Video Gencalesthio/OpenMontage66k—~3kAutomated safety check: PassAGPL-3.0
AI Video Gencalesthio/OpenMontage66k—~2.8kAutomated safety check: PassAGPL-3.0

Similar skills

  • Higgsfield

    OSideMedia/higgsfield-ai-prompt-skill

    A skill your agent uses whenever the user asks anything about Higgsfield AI — writing or refining video/image prompts, choosing a model (Kling, Veo, Wan, Seedance, Minimax Hailuo, DoP, Soul, Nano…

    713 GitHub stars~9.1k tokensUpdated 13 days ago
    Media & CreativeAuto-check passed
  • Video

    Nexus-JPF/note-companion

    When the user wants to create, generate, or produce video content using AI tools or programmatic frameworks.

    870 GitHub starsUsed in 3 repos~3.6k tokens
    Media & CreativeAuto-check passed
  • Video Prompting

    Square-Zero-Labs/video-prompting-skill

    Draft and refine prompts for video generation models (including text-to-video, image/keyframe-to-video, and reference-driven generation), and create character-sheet prompts for image models when the…

    182 GitHub stars~1.9k tokensUpdated 1 mo ago
    Media & CreativeAuto-check passed
  • AI Video Gen

    calesthio/OpenMontage

    Generate AI videos from text prompts using multiple provider gateways.

    66k GitHub stars~3k tokensUpdated 7 days ago
    Media & CreativeAuto-check passed
  • AI Video Gen

    calesthio/OpenMontage

    Generate AI videos from text prompts using multiple provider gateways.

    66k GitHub stars~2.8k tokensUpdated 7 days ago
    Media & CreativeAuto-check passed
  • Atlas Cloud

    calesthio/OpenMontage

    Generate or edit images and videos through the Atlas Cloud gateway.

    66k GitHub stars~1.2k tokensUpdated 7 days ago
    Media & CreativeAuto-check passed

More from pexoai/pexo-skills

All 18 skills in this repo
  • AI Video Generation

    pexoai/pexo-skills

    Generate AI video from any input — text, image, or script — with Pexo.

    804 GitHub stars~1.3k tokensUpdated 1 mo ago
    Auto-check passed
  • Explainer Video

    pexoai/pexo-skills

    Create an explainer video with narration using Pexo. An agent skill from pexoai/pexo-skills.

    804 GitHub stars~1.3k tokensUpdated 1 mo ago
    Auto-check passed
  • Founder Video

    pexoai/pexo-skills

    Make a founder video with Pexo — built for solo founders and small teams.

    804 GitHub stars~1.4k tokensUpdated 1 mo ago
    Auto-check passed
  • Image To Video

    pexoai/pexo-skills

    Animate a still image into a finished, moving video with Pexo.

    804 GitHub stars~1.3k tokensUpdated 1 mo ago
    Auto-check passed
  • Launch Video

    pexoai/pexo-skills

    Make a launch video for your startup or product with Pexo. An agent skill from pexoai/pexo-skills.

    804 GitHub stars~1.4k tokensUpdated 1 mo ago
    Auto-check passed
  • Make A Video

    pexoai/pexo-skills

    Make a complete video from a simple idea with Pexo. An agent skill from pexoai/pexo-skills.

    804 GitHub stars~1.3k tokensUpdated 1 mo ago
    Auto-check passed

Works with

Questions about Videoagent Video Studio

What does Videoagent Video Studio do?

Generate short AI videos from text or images — text-to-video, image-to-video, and reference-based generation — with zero API key setup. Videoagent Video Studio is an agent skill from pexoai/pexo-skills. Generate short AI videos from text or images — text-to-video, image-to-video, and reference-based generation — with zero API key setup.

When should I use Videoagent Video Studio?

Videoagent Video Studio fits situations like: the user wants to create a video clip; animate an image; generate video from a description.

How do I install Videoagent Video Studio in Claude Code?

Run `npx skills add pexoai/pexo-skills --skill videoagent-video-studio -a claude-code`. Or copy the skill folder (skills/videoagent-video-studio in pexoai/pexo-skills) into .claude/skills/videoagent-video-studio in your project. Claude Code loads it when a task matches its description.

How do I install Videoagent Video Studio in Codex?

Run `npx skills add pexoai/pexo-skills --skill videoagent-video-studio -a codex`. Or copy the skill folder (skills/videoagent-video-studio in pexoai/pexo-skills) into .agents/skills/videoagent-video-studio in your project. Codex loads it when a task matches its description.

Can I use Videoagent Video Studio in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add pexoai/pexo-skills --skill videoagent-video-studio -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/videoagent-video-studio, .gemini/skills/videoagent-video-studio, .github/skills/videoagent-video-studio and .opencode/skills/videoagent-video-studio in your project.

What does Videoagent Video Studio need to run?

Going by SKILL.md and its folder, Videoagent Video Studio needs JavaScript for the scripts in its folder, the command-line tools its instructions call (node) and credentials named VIDEO_STUDIO_TOKEN. Our summary lists: Node.js; A credential in VIDEO_STUDIO_TOKEN.

Does Videoagent Video Studio access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is Videoagent Video Studio safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. The check reads SKILL.md only: the scripts in the folder are not scanned, so read them before running anything.

What licence does Videoagent Video Studio use?

Videoagent Video Studio is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Videoagent Video Studio use?

About 1.6k tokens (SKILL.md is roughly 6.4k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full. Its references folder adds about 2.7k tokens, read only when the agent opens those files.

What are the alternatives to Videoagent Video Studio?

Skills that share tags, products or a category with Videoagent Video Studio: Higgsfield (OSideMedia/higgsfield-ai-prompt-skill, 713 stars), Video (Nexus-JPF/note-companion, 870 stars), Video Prompting (Square-Zero-Labs/video-prompting-skill, 182 stars) and AI Video Gen (calesthio/OpenMontage, 66k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Videoagent Video Studio?

pexoai (a GitHub user) maintains it in pexoai/pexo-skills, which has 804 GitHub stars. The repository holds 18 skills in this directory. The repository was last updated on August 20, 2026.

Source: pexoai/pexo-skills on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.