A skill your agent uses to write great prompts for Google Veo (Veo 3.1) to generate or animate video for social media — the video-prompt-craft mini-skill, the video counterpart to nano-banana.

MITAuto-check passedMedia & Creative

Install Veo 3

skills CLI
$ npx skills add social-media-skills/skills --skill veo-3 -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install social-media-skills/skills veo-3 --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/social-media-skills/skills.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/veo-3 .claude/skills/veo-3 && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
veo-3
GitHub stars
134
Token cost
~1.8k tokens
SKILL.md length
687 words
Files
6 (incl. references)
Skills in repo
106
Repo updated
First seen
Licence
MIT

At a glance

A skill your agent uses to write great prompts for Google Veo (Veo 3.1) to generate or animate video for social media — the video-prompt-craft mini-skill, the video counterpart to nano-banana.

  • Works in 5 steps: Read the brand + the job → Direct the shot (natural language, not… → Use the superpower: audio → …
  • Write great prompts for Google Veo (Veo 3.
  • SKILL.md covers Reach for Veo when… (match the…, Step 0 — Read the brand + the…, Step 1 — Direct the shot… and Step 2 — Use the superpower:…, plus 6 more sections
  • Instructions only: no scripts, shell commands, URLs or credentials in SKILL.md

What it does

Veo 3 is an agent skill from social-media-skills/skills. Use to write great prompts for Google Veo (Veo 3.1) to generate or animate video for social media — the video-prompt-craft mini-skill, the video counterpart to nano-banana. Run when the user wants a Veo / AI video prompt, a short video clip for a post (b-roll, product-in-motion, hook visual, spokesperson/UGC clip, ad), to animate a still image into video, or a vertical Reel/TikTok/Short clip. Reads brand-profile for brand style. Veo's standout is native synchronized audio in one pass, plus image-to-video and…

Its SKILL.md is about 1.8k tokens, which your agent loads only when the skill is triggered. The skill folder holds 7 other files, including reference files (for example `evals/evals.json`, `references/audio-and-camera.md` and `references/examples.md`).

It sits in Media & Creative, covering AI video generation. It works with Google Veo, Google Gemini and TikTok. The repository describes itself as: 106 social media skills for AI agents - strategy, writing, video, design, platform growth, publishing, and analytics. Works with Claude, Cursor, OpenClaw, Hermes & 40+ agents. The licence is MIT.

When your agent uses it

  • Write great prompts for Google Veo (Veo 3.
  • Animate video for social media — the video-prompt-craft mini-skill
  • The video counterpart to nano-banana
  • Wants a Veo / AI video prompt

Example prompts

  • “/veo-3”

Workflow steps

5 steps, taken from the step headings in SKILL.md.

  1. Read the brand + the job
  2. Direct the shot (natural language, not quality spam)
  3. Use the superpower: audio
  4. Inputs, length, recipes
  5. Iterate cheaply, then verify, disclose, ship

What it can do on your machine

Read from SKILL.md and the folder at commit 6e30eeb. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    No scripts in the folder and no shell commands in SKILL.md.

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Veo 3 loads about 1.8k tokens when it runs, and up to ~5k if it reads all its reference files. Until then it costs about 251 tokens; SKILL.md has 687 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~251
When it runs · the whole SKILL.md, loaded when a task matches
~1.8k
With references · SKILL.md plus every file in references/, read only if the agent opens them
~5k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from social-media-skills/skills at commit 6e30eeb, republished under its MIT licence (© social-media-skills). 687 words, ~1,753 tokens.

Download SKILL.mdSave it as .claude/skills/veo-3/SKILL.md (or your agent's skills folder). This skill also uses 5 other files; get the full folder from GitHub.
name
veo-3
description
Use to write great prompts for Google Veo (Veo 3.1) to generate or animate video for social media — the video-prompt-craft mini-skill, the video counterpart to nano-banana. Run when the user wants a Veo / AI video prompt, a short video clip for a post (b-roll, product-in-motion, hook visual, spokesperson/UGC clip, ad), to animate a still image into video, or a vertical Reel/TikTok/Short clip. Reads brand-profile for brand style. Veo's standout is native synchronized audio in one pass, plus image-to-video and native 9:16 vertical. Teaches the prompt anatomy, audio prompting, the image-to-video pipeline, and the 8-second constraint + extension/stitching. Honest: iterate cheap then finalize, disclose AI video (SynthID watermark), never generate real identifiable people or copyrighted IP. This is the prompt-craft layer — the API/connection and the generate -> upload to WoopSocial Media -> attach flow live in tools/integrations/veo.md; the consuming pack is reels-script's veo-prompt-pack.
metadata.version
1.0.0
license
MIT

Veo (video prompt craft)

Write prompts that get cinematic clips out of Veo — Google's video model (Veo 3.1; Fast/Lite tiers for cheap iteration). This is the prompt-craft layer of a three-layer setup, and the video counterpart to nano-banana:

  • Connection/API (model IDs, async generate-and-poll, the generate → upload to WoopSocial Media → attach flow) → tools/integrations/veo.md.
  • Prompt craft (this skill) → how to direct the right clip well.
  • In-skill application → reels-script's veo-prompt-pack.

Fast-moving area — re-verify model names/specs quarterly.

Reach for Veo when… (match the job)

Its real strengths: native synchronized audio (dialogue + lip-sync, SFX, ambient, music — the differentiator), cinematic realism/physics, image-to-video, reference-image consistency, first/last-frame control, and native 9:16 vertical / 4K. Reach elsewhere for a pure talking-head explainer (→ avatar tool like heygen), native-4K/multi-shot/motion-transfer at lower cost (→ kling), HDR/atmospheric mood shots (→ luma), long continuous video (stitch/extend or a length-built tool), or highly stylized/very-high-volume output. Don't default it to every job. (Details: references/when-and-how-to-prompt.md.)

Step 0 — Read the brand + the job

Load brand-profile.md (visual style, palette, tone). Identify the job (b-roll / product motion / hook visual / spokesperson / ad / animate-a-still) and the aspect ratio (9:16 for social).

Step 1 — Direct the shot (natural language, not quality spam)

Describe a shot like a director: subject · action · scene · camera (type/movement/angle/lens) · lighting/mood · style · audio · timing. One clear motion per clip. Set 9:16. Ground in the brand. See references/when-and-how-to-prompt.md.

Step 2 — Use the superpower: audio

Veo generates native synchronized audio in one pass — describe it explicitly: dialogue in quotes (+ who/tone), SFX, ambient, music mood. Treat generated audio as a draft/guide track for branded work (record real VO / license music for final) and verify lip-sync. Direct the camera/motion for the cinematic feel. See references/audio-and-camera.md.

Step 3 — Inputs, length, recipes

  • Image-to-video: animate a still (e.g. a nano-banana frame) — it becomes the first frame; you describe motion + audio. Use reference images / first-and-last frame for consistency/control.
  • Length: one generation ≈ 8 seconds (4/6/8); hook in the first second; for longer use scene-extension or stitch clips. Plan short beats.
  • Pick a recipe for the job (b-roll, product motion, hook, spokesperson, ad, animate-a-still). See references/inputs-length-and-recipes.md.

Step 4 — Iterate cheaply, then verify, disclose, ship

  • Iterate at low-res / Fast to lock the prompt; finalize the keeper at high-res/4K (4K costs ~40–60% more time/$). Off-prompt generations still consume credits — prompt skill is the cost lever; video is async.
  • Verify every clip (artifacts, lip-sync, physics) before publishing.
  • Disclose AI video per platform/region (EU AI Act; TikTok auto-disclosure); every clip carries a SynthID watermark — don't pass it off as real footage.
  • Ship: generate per the integration guide → upload to WoopSocial Media → attach via scheduling-and-queue. WoopSocial doesn't generate video.
Show full SKILL.md (266 more words)Show less

Quality bar — self-check

  • Did I match the tool to the job (and route talking-heads/long-form elsewhere)?
  • Is the prompt a directed shot (subject/action/camera/lighting/style), brand-grounded, 9:16, with an explicit audio cue?
  • Did I plan for 8-second clips (hook first) and iterate cheap → finalize?
  • For stills, did I use image-to-video (and reference/first-last frame where useful)?
  • Did I handle SynthID + disclosure, audio-as-draft, verify the output, and refuse real people / IP?
  • Did I point to tools/integrations/veo.md for the API + WoopSocial flow (no claim WoopSocial generates video)?

Edge cases & pushback

  • Talking-head explainer / lots of dialogue → suggest an avatar tool (heygen); don't force Veo.
  • "Make a 40-second video" → ~8s per gen; scene-extend/stitch; plan short beats.
  • "Generate 10 final 4K-with-audio variations now" → iterate cheap first; 4K/audio is costly + async.
  • "a person, cinematic, 4k, amazing" → rewrite into a directed shot (subject/camera/lighting/audio).
  • Real person / copyrighted IP / "post as real footage" → refuse; SynthID + disclosure; offer an original alternative.
  • "Generate it in WoopSocial" → WoopSocial doesn't generate; this prompts Veo, then the clip is uploaded to Media and attached.
  • tools/integrations/veo.md — API/model IDs, async generate-and-poll, pricing, the upload-to-WoopSocial flow.
  • reels-script (veo-prompt-pack) — the consuming skill; nano-banana — the image sibling + image-to-video source.
  • brand-profile — the visual brand; hook-writer — the in-clip hook/line; heygen — avatar/talking-head alternative.
  • ai-video — the router above this skill; kling (4K/multi-shot/motion) and luma (HDR/mood, silent) — generative siblings.
  • scheduling-and-queue — attach the video to a post and publish.

References

  • references/when-and-how-to-prompt.md — when to reach for Veo vs other tools, and the shot/prompt anatomy.
  • references/audio-and-camera.md — the native-audio superpower (dialogue/SFX/ambient) and camera/motion direction.
  • references/inputs-length-and-recipes.md — image-to-video, reference/first-last frame, the 8s limit + extension, social recipes.
  • references/examples.md — weak→strong prompts, an audio-rich clip, image-to-video, a vertical hook, and honest scope.

© social-media-skills, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

SKILL.md and 5 other files (references) in skills/veo-3 of social-media-skills/skills.

  • SKILL.md
  • evals/evals.json
  • references/audio-and-camera.md
  • references/examples.md
  • references/inputs-length-and-recipes.md
  • references/when-and-how-to-prompt.md

Open the folder on GitHubat commit 6e30eeb

Compare with similar skills

Veo 3 next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Veo 3 compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Veo 3 this skillsocial-media-skills/skills134—~1.8kAutomated safety check: PassMIT
AI Media GeneratorHao0321/ai-media-generator260—~5.6kAutomated safety check: PassMIT
Arcads External APIkrusemediallc/arcads-claude-code1.6k—~8.7kAutomated safety check: NotesMIT
Fal AI Mediaaffaan-m/ECC277k4 repos~1.9kAutomated safety check: PassMIT
Fal AI Mediaaffaan-m/ECC277k2 repos~1.2kAutomated safety check: PassMIT
Fal AI Mediaaffaan-m/ECC276k—~1.4kAutomated safety check: PassMIT

Similar skills

  • AI Media Generator

    Hao0321/ai-media-generator

    為使用者產生高品質的 AI 生圖、生影片、生音樂提示詞,並在需要時透過瀏覽器自動化實際送到目標平台。涵蓋 OiiOii、Kling 3.0/O-series、Seedance 2.0/2.5、Suno v5.5、Seedream 5.0/4.0、Vidu Q3、Midjourney V8.1、Flux 1.1 Pro / Kontext、Runway Gen-4.5 /…

    260 GitHub stars~5.6k tokensUpdated 2 mo ago
    Media & CreativeAuto-check passed
  • Arcads External API

    krusemediallc/arcads-claude-code

    Creates and retrieves AI video and image-related assets via the Arcads external API (Seedance 2.0, Sora 2, Veo 3.1, Kling, Grok Video, Nano Banana, b-roll, scene, script/actor flows).

    1.6k GitHub stars~8.7k tokensUpdated 18 days ago
    Media & CreativeAuto-check: notes
  • Fal AI Media

    affaan-m/ECC

    Unified media generation via fal.ai MCP — image, video, and audio.

    277k GitHub starsUsed in 4 repos~1.9k tokens
    Media & CreativeAuto-check passed
  • Fal AI Media

    affaan-m/ECC

    通过 fal.ai MCP 实现统一的媒体生成——图像、视频和音频。涵盖文本到图像(Nano Banana)、文本/图像到视频(Seedance、Kling、Veo 3)、文本到语音(CSM-1B),以及视频到音频(ThinkSound)。当用户想要使用 AI 生成图像、视频或音频时使用。

    277k GitHub starsUsed in 2 repos~1.2k tokens
    Media & CreativeAuto-check passed
  • Fal AI Media

    affaan-m/ECC

    fal.ai MCPによる統合メディア生成(画像、動画、音声)。テキストから画像(Nano Banana)、テキスト/画像から動画(Seedance、Kling、Veo 3)、テキストから音声(CSM-1B)、動画から音声(ThinkSound)をカバーします。ユーザーがAIで画像、動画、音声を生成したい場合に使用します。

    276k GitHub stars~1.4k tokensUpdated yesterday
    Media & CreativeAuto-check passed
  • Higgsfield Models

    OSideMedia/higgsfield-ai-prompt-skill

    A skill your agent uses when the user asks which model to use, wants to compare models, or needs guidance on selecting between Kling, Wan (incl.

    713 GitHub stars~7k tokensUpdated 14 days ago
    Media & CreativeAuto-check passed

More from social-media-skills/skills

All 106 skills in this repo
  • AI Image Editing

    social-media-skills/skills

    The AI image-editing router — inpainting/object removal, background removal, upscaling, outpainting, old-photo restoration, and retouch, routed task-first to the right engine.

    134 GitHub stars~2.1k tokensUpdated 9 days ago
    Auto-check passed
  • AI Music And Sound

    social-media-skills/skills

    The AI music + sound-design skill for social -- original/licensed audio beds and sound design for Reels/TikToks/Shorts/videos.

    134 GitHub stars~1.9k tokensUpdated 9 days ago
    Auto-check passed
  • AI Search Optimization

    social-media-skills/skills

    A skill your agent uses to get a brand and its content CITED and RECOMMENDED by AI answer engines — the GEO (Generative Engine Optimization) / AI-search-visibility skill.

    134 GitHub stars~2k tokensUpdated 9 days ago
    Auto-check passed
  • AI Video

    social-media-skills/skills

    The model-agnostic AI-video router and brief — the counterpart to image-prompt.

    134 GitHub stars~1.3k tokensUpdated 9 days ago
    Auto-check passed
  • AI Voiceover

    social-media-skills/skills

    The AI narration / voiceover mini-skill (ElevenLabs-led). An agent skill from social-media-skills/skills.

    134 GitHub stars~1.1k tokensUpdated 9 days ago
    Auto-check passed
  • Analytics And Reporting

    social-media-skills/skills

    Social media analytics and reporting — read native platform data honestly and turn it into next actions.

    134 GitHub stars~1.4k tokensUpdated 9 days ago
    Auto-check passed

Questions about Veo 3

What does Veo 3 do?

A skill your agent uses to write great prompts for Google Veo (Veo 3.1) to generate or animate video for social media — the video-prompt-craft mini-skill, the video counterpart to nano-banana. Veo 3 is an agent skill from social-media-skills/skills.1) to generate or animate video for social media — the video-prompt-craft mini-skill, the video counterpart to nano-banana.

When should I use Veo 3?

Veo 3 fits situations like: write great prompts for Google Veo (Veo 3; animate video for social media — the video-prompt-craft mini-skill; the video counterpart to nano-banana; wants a Veo / AI video prompt.

How do I install Veo 3 in Claude Code?

Run `npx skills add social-media-skills/skills --skill veo-3 -a claude-code`. Or copy the skill folder (skills/veo-3 in social-media-skills/skills) into .claude/skills/veo-3 in your project. Claude Code loads it when a task matches its description.

How do I install Veo 3 in Codex?

Run `npx skills add social-media-skills/skills --skill veo-3 -a codex`. Or copy the skill folder (skills/veo-3 in social-media-skills/skills) into .agents/skills/veo-3 in your project. Codex loads it when a task matches its description.

Can I use Veo 3 in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add social-media-skills/skills --skill veo-3 -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/veo-3, .gemini/skills/veo-3, .github/skills/veo-3 and .opencode/skills/veo-3 in your project.

What does Veo 3 need to run?

SKILL.md names no scripts, command-line tools or credentials: Veo 3 is instructions for the agent only.

Does Veo 3 access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is Veo 3 safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Veo 3 use?

Veo 3 is published under the MIT licence (declared in SKILL.md). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Veo 3 use?

About 1.8k tokens (SKILL.md is roughly 7k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full. Its references folder adds about 3.2k tokens, read only when the agent opens those files.

What are the alternatives to Veo 3?

Skills that share tags, products or a category with Veo 3: AI Media Generator (Hao0321/ai-media-generator, 260 stars), Arcads External API (krusemediallc/arcads-claude-code, 1.6k stars), Fal AI Media (affaan-m/ECC, 277k stars) and Fal AI Media (affaan-m/ECC, 277k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Veo 3?

social-media-skills (a GitHub organization) maintains it in social-media-skills/skills, which has 134 GitHub stars. The repository holds 106 skills in this directory. The repository was last updated on October 1, 2026.

Source: social-media-skills/skills on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.