Agent skill

Scenario Veo

by scenario-labs in scenario-labs/skills

A skill your agent uses when generating or extending video with Google Veo models on Scenario via MCP: text-to-video, image-to-video from a first frame, first and last frame transitions…

MITAuto-check passedMedia & Creative

Install Scenario Veo

skills CLI
$ npx skills add scenario-labs/skills --skill scenario-veo -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install scenario-labs/skills scenario-veo --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/scenario-labs/skills.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/scenario-veo .claude/skills/scenario-veo && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
scenario-veo
GitHub stars
946
Token cost
~1.4k tokens
SKILL.md length
691 words
Files
1
Skills in repo
146
Repo updated
First seen
Licence
MIT

At a glance

A skill your agent uses when generating or extending video with Google Veo models on Scenario via MCP: text-to-video, image-to-video from a first frame, first and last frame transitions…

  • Works in 6 steps: search with target="models",… → model_schema_get on the chosen id:… → upload_asset two character stills (see… → …
  • Extending video with Google Veo models on Scenario via MCP: text-to-video
  • SKILL.md covers Overview, Quick reference, The soundstage is in the prompt and Prompt one continuous shot, plus 2 more sections
  • Calls npx

What it does

Scenario Veo is an agent skill from scenario-labs/skills. Use when generating or extending video with Google Veo models on Scenario via MCP: text-to-video, image-to-video from a first frame, first and last frame transitions, reference-to-video (R2V) with asset or style reference images, native audio with dialogue, sound effects, and ambience, negative prompts, seeded reruns, extending a 16:9 clip, or choosing a quality, fast, or lite tier. Keywords: Veo 3.1, Veo 3.1 Fast, Veo 3.1 Lite, Extend Video, Google, T2V, I2V, V2V, 720p, 1080p.

Its SKILL.md is about 1.4k tokens, which your agent loads only when the skill is triggered. It is a single SKILL.md file with no bundled scripts.

It sits in Media & Creative, covering AI video generation. It works with Google Veo and Model Context Protocol. The repository describes itself as: Get production-ready images, video, audio, and 3D from any AI agent: skills that pick the right model, price before spending, and keep characters and brands consistent through… The licence is MIT.

When your agent uses it

  • Extending video with Google Veo models on Scenario via MCP: text-to-video
  • Image-to-video from a first frame
  • First and last frame transitions
  • Reference-to-video (R2V) with asset

Example prompts

  • “/scenario-veo”

Requirements

  • Node.js

Workflow steps

6 steps, taken from the first numbered list in SKILL.md.

  1. search with target="models", query="veo", public=true. Prefer the newest non-deprecated hits, e.g. model_veo3-1 and model_veo3-1-fast…
  2. model_schema_get on the chosen id: confirm which reference inputs exist before writing the payload.
  3. upload_asset two character stills (see the scenario skill) to get asset ids.
  4. model_run with that model_id, dry_run=true, and parameters={"prompt": "Medium shot, the knight lowers her visor and says, \"Hold the…
  5. Repeat model_run with wait=false, then jobs_wait with the returned job id, re-called with pending_job_ids on timeout, never a second…
  6. asset_display the output and review with sound. To iterate on wording alone, hold an explicit seed fixed across reruns.

What it can do on your machine

Read from SKILL.md and the folder at commit f6f8ab7. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Shell commands in SKILL.md call:

    • npx

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md. Its commands use npx, which can reach the network depending on how they are called.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Scenario Veo loads about 1.4k tokens when it runs. Until then it costs about 124 tokens; SKILL.md has 691 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~124
When it runs · the whole SKILL.md, loaded when a task matches
~1.4k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from scenario-labs/skills at commit f6f8ab7, republished under its MIT licence (© scenario-labs). 691 words, ~1,419 tokens.

Download SKILL.mdSave it as .claude/skills/scenario-veo/SKILL.md (or your agent's skills folder).
name
scenario-veo
description
Use when generating or extending video with Google Veo models on Scenario via MCP: text-to-video, image-to-video from a first frame, first and last frame transitions, reference-to-video (R2V) with asset or style reference images, native audio with dialogue, sound effects, and ambience, negative prompts, seeded reruns, extending a 16:9 clip, or choosing a quality, fast, or lite tier. Keywords: Veo 3.1, Veo 3.1 Fast, Veo 3.1 Lite, Extend Video, Google, T2V, I2V, V2V, 720p, 1080p.
license
MIT

Scenario Veo Video

Overview

Veo, Google's video family on Scenario, ships four members at authoring time: Veo 3.1 (full quality, reference images with a style switch), Veo 3.1 Fast (same modes, quicker, no switch), Veo 3.1 Lite (frame anchors only, cheapest), and Veo 3.1 Extend Video (continues a clip). Discover them with search and treat model_schema_get as the contract: members agree on parameter names and disagree on which inputs exist at all.

Connection and the core loop: see the scenario skill in this repo; model-agnostic video work: the scenario-video skill. If a sibling skill named here is missing from your available skills, ask the user to install it (npx skills add scenario-labs/skills --skill <name>); unattended, proceed from tool schemas and flag the gap.

Quick reference

Mode follows from the inputs (names from the live schema):

ModeInputsBehavior
TextpromptaspectRatio 16:9 or 9:16
First frameimage (+ prompt)opens on that frame; a mismatched image may be cropped
Transitionimage + lastFrameImagebridges the stills; prompt one continuous move
ReferencereferenceImages (1 to 3)subject or style consistency; duration locked to 8
Extendvideo (Extend member)continues the clip; 16:9 source, short side 720p or 1080p

image and referenceImages are mutually exclusive. referenceImagesType exists on full Veo 3.1 alone and reads the whole array one way: ASSET (default) carries subjects, objects, and scenes; STYLE carries palette, lighting, and texture. Fast takes references without the switch; Lite takes none. Shared across the three generators at authoring time: negativePrompt, resolution (720p, 1080p), duration (4, 6, or 8 seconds), and seed. generateAudio is required on every member and moves the price. The Extend member takes only prompt, video, generateAudio, and seed: no duration, resolution, or ratio controls. At authoring time price spanned roughly five times between Lite and full 3.1 for one image-to-video job, and one extension cost more than a fresh full generation, so aim for the shot in one 8 second pass and dry_run the same payload on two members before a batch.

The soundstage is in the prompt

With generateAudio: true, Veo renders synchronized dialogue, effects, and ambience, and it takes audio direction literally. Put spoken lines in quotation marks, prefix effects with SFX: and room tone with Ambient noise:. A prompt with no audio direction still gets a soundtrack, just not the one you meant. When a clip needs silence for later scoring, set generateAudio: false rather than prompting for quiet.

Show full SKILL.md (290 more words)Show less

Prompt one continuous shot

Write present-tense prose ordered as cinematography, subject, action, context, style: camera and shot scale first, one primary arc, roughly 90 to 250 words. For several beats in one clip, timestamp lines work ([00:00-00:02] Medium shot...), spans fitting inside the chosen duration.

Worked example: a dialogue shot from character references

  1. search with target="models", query="veo", public=true. Prefer the newest non-deprecated hits, e.g. model_veo3-1 and model_veo3-1-fast (live hits at authoring time: re-discover each session).
  2. model_schema_get on the chosen id: confirm which reference inputs exist before writing the payload.
  3. upload_asset two character stills (see the scenario skill) to get asset ids.
  4. model_run with that model_id, dry_run=true, and parameters={"prompt": "Medium shot, the knight lowers her visor and says, \"Hold the line.\" Torchlight flickers on wet stone. Ambient noise: distant thunder. SFX: metal visor clank.", "referenceImages": ["asset_a", "asset_b"], "referenceImagesType": "ASSET", "duration": 8, "resolution": "1080p", "aspectRatio": "16:9", "generateAudio": true}. Estimate the same job on the Fast id (drop referenceImagesType, absent there) and compare.
  5. Repeat model_run with wait=false, then jobs_wait with the returned job id, re-called with pending_job_ids on timeout, never a second model_run.
  6. asset_display the output and review with sound. To iterate on wording alone, hold an explicit seed fixed across reruns.

Common mistakes

  • Combining image with referenceImages: mutually exclusive; pick the opening state or consistency.
  • Any duration but 8 with referenceImages: reference mode supports only 8 seconds.
  • Sending referenceImages to Lite or referenceImagesType to Fast: the schema decides which inputs exist.
  • Omitting generateAudio: it is required on every member; pass it explicitly.
  • Extending a 9:16 clip: Extend takes 16:9 sources with a 720p or 1080p short side only.
  • Writing the prompt as a list of negatives: describe the wanted scene, reserve negativePrompt for what to discourage.

© scenario-labs, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

Just SKILL.md in skills/scenario-veo of scenario-labs/skills.

Open the folder on GitHubat commit f6f8ab7

Compare with similar skills

Scenario Veo next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Scenario Veo compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Scenario Veo this skillscenario-labs/skills946—~1.4kAutomated safety check: PassMIT
Fal AI Mediaaffaan-m/ECC277k4 repos~1.9kAutomated safety check: PassMIT
Fal AI Mediaaffaan-m/ECC277k2 repos~1.2kAutomated safety check: PassMIT
Fal AI Mediaaffaan-m/ECC276k—~1.4kAutomated safety check: PassMIT
Avatar Videocalesthio/OpenMontage66k—~1.6kAutomated safety check: PassAGPL-3.0
Create Videocalesthio/OpenMontage66k—~1.3kAutomated safety check: PassAGPL-3.0

Similar skills

  • Fal AI Media

    affaan-m/ECC

    Unified media generation via fal.ai MCP — image, video, and audio.

    277k GitHub starsUsed in 4 repos~1.9k tokens
    Media & CreativeAuto-check passed
  • Fal AI Media

    affaan-m/ECC

    通过 fal.ai MCP 实现统一的媒体生成——图像、视频和音频。涵盖文本到图像(Nano Banana)、文本/图像到视频(Seedance、Kling、Veo 3)、文本到语音(CSM-1B),以及视频到音频(ThinkSound)。当用户想要使用 AI 生成图像、视频或音频时使用。

    277k GitHub starsUsed in 2 repos~1.2k tokens
    Media & CreativeAuto-check passed
  • Fal AI Media

    affaan-m/ECC

    fal.ai MCPによる統合メディア生成(画像、動画、音声)。テキストから画像(Nano Banana)、テキスト/画像から動画(Seedance、Kling、Veo 3)、テキストから音声(CSM-1B)、動画から音声(ThinkSound)をカバーします。ユーザーがAIで画像、動画、音声を生成したい場合に使用します。

    276k GitHub stars~1.4k tokensUpdated yesterday
    Media & CreativeAuto-check passed
  • Avatar Video

    calesthio/OpenMontage

    Create AI avatar videos with precise control over avatars, voices, scripts, scenes, and backgrounds using HeyGen's v2 API.

    66k GitHub stars~1.6k tokensUpdated 7 days ago
    Media & CreativeAuto-check passed
  • Create Video

    calesthio/OpenMontage

    Create videos from a text prompt using HeyGen's Video Agent.

    66k GitHub stars~1.3k tokensUpdated 7 days ago
    Media & CreativeAuto-check passed
  • Cassette Video Edit

    Cassette-Editor/oh-my-cassette

    Edit, trim, cut, caption, subtitle, reframe, combine, add background music to, or export video, audio, and image files through Cassette.

    119 GitHub starsUsed in 1 repo~3.4k tokens
    Media & CreativeAuto-check passed

More from scenario-labs/skills

All 146 skills in this repo
  • Scenario Blender Grease Pencil

    scenario-labs/skills

    A skill your agent uses when drawing or animating with Grease Pencil in Blender 5.x from Python: 2D or 2.5D illustration, frame-by-frame animation, a cutout or part-based 2D character, strokes with…

    946 GitHub stars~4.5k tokensUpdated yesterday
    Auto-check passed
  • Scenario Blender Hair

    scenario-labs/skills

    A skill your agent uses when grooming hair or fur in Blender with hair curves, such as a character hairstyle, animal fur, procedural fur in geometry nodes, or hair cards and mesh hair for games.

    946 GitHub stars~4.7k tokensUpdated yesterday
    Auto-check passed
  • A skill your agent uses when lighting, rendering or compositing in Blender: light a character, product or hero shot, interior at dusk or night, three-point or motivated lighting, sun and sky, HDRI…

    946 GitHub stars~5k tokensUpdated yesterday
    Auto-check passed
  • Scenario Chatgpt Pet Create

    scenario-labs/skills

    A skill your agent uses when creating a ChatGPT pet or Codex pet with Scenario: hatching an animated companion from a text idea, a character, mascot or brand cue, or reference photos and art; making…

    946 GitHub stars~3.6k tokensUpdated yesterday
    Auto-check passed
  • Scenario Godot Animation

    scenario-labs/skills

    A skill your agent uses when animating characters or scenes in Godot 4.7: AnimationPlayer clips and RESET, AnimationTree state machines and blend spaces built in code, Mixamo or glTF import, loop…

    946 GitHub stars~4.7k tokensUpdated yesterday
    Auto-check passed
  • Scenario Godot Audio

    scenario-labs/skills

    A skill your agent uses when adding or fixing sound in Godot 4.7: audio buses and effects, volume sliders, 'too many sounds', combat audio with hundreds of enemies, sounds clipping or distorting, 3D…

    946 GitHub stars~4.7k tokensUpdated yesterday
    Auto-check passed

Questions about Scenario Veo

What does Scenario Veo do?

A skill your agent uses when generating or extending video with Google Veo models on Scenario via MCP: text-to-video, image-to-video from a first frame, first and last frame transitions…. Scenario Veo is an agent skill from scenario-labs/skills. Use when generating or extending video with Google Veo models on Scenario via MCP: text-to-video, image-to-video from a first frame, first and last frame transitions, reference-to-video (R2V) with asset or style reference images, native audio with dialogue, sound effects, and ambience, negative prompts, seeded reruns, extending a 16:9 clip, or choosing a quality, fast, or lite tier.

When should I use Scenario Veo?

Scenario Veo fits situations like: extending video with Google Veo models on Scenario via MCP: text-to-video; image-to-video from a first frame; first and last frame transitions; reference-to-video (R2V) with asset.

How do I install Scenario Veo in Claude Code?

Run `npx skills add scenario-labs/skills --skill scenario-veo -a claude-code`. Or copy the skill folder (skills/scenario-veo in scenario-labs/skills) into .claude/skills/scenario-veo in your project. Claude Code loads it when a task matches its description.

How do I install Scenario Veo in Codex?

Run `npx skills add scenario-labs/skills --skill scenario-veo -a codex`. Or copy the skill folder (skills/scenario-veo in scenario-labs/skills) into .agents/skills/scenario-veo in your project. Codex loads it when a task matches its description.

Can I use Scenario Veo in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add scenario-labs/skills --skill scenario-veo -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/scenario-veo, .gemini/skills/scenario-veo, .github/skills/scenario-veo and .opencode/skills/scenario-veo in your project.

What does Scenario Veo need to run?

Going by SKILL.md and its folder, Scenario Veo needs the command-line tools its instructions call (npx). Our summary lists: Node.js.

Does Scenario Veo access the network?

SKILL.md contains no URLs. Its commands use npx, which can reach the network depending on how they are called. This is read from the text; nothing was executed.

Is Scenario Veo safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Scenario Veo use?

Scenario Veo is published under the MIT licence (declared in SKILL.md). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Scenario Veo use?

About 1.4k tokens (SKILL.md is roughly 5.7k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Scenario Veo?

Skills that share tags, products or a category with Scenario Veo: Fal AI Media (affaan-m/ECC, 277k stars), Fal AI Media (affaan-m/ECC, 277k stars), Fal AI Media (affaan-m/ECC, 276k stars) and Avatar Video (calesthio/OpenMontage, 66k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Scenario Veo?

scenario-labs (a GitHub organization) maintains it in scenario-labs/skills, which has 946 GitHub stars. The repository holds 146 skills in this directory. The repository was last updated on October 10, 2026.

Source: scenario-labs/skills on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.