Agent skill

Scenario Vidu

by scenario-labs in scenario-labs/skills

A skill your agent uses when generating video with Vidu models on Scenario via MCP: text-to-video, image-to-video from a single still, start and end frame interpolation, reference-to-video keeping…

MITAuto-check passedMedia & Creative

Install Scenario Vidu

skills CLI
$ npx skills add scenario-labs/skills --skill scenario-vidu -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install scenario-labs/skills scenario-vidu --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/scenario-labs/skills.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/scenario-vidu .claude/skills/scenario-vidu && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
scenario-vidu
GitHub stars
946
Token cost
~1.5k tokens
SKILL.md length
785 words
Files
1
Skills in repo
146
Repo updated
First seen
Licence
MIT

At a glance

A skill your agent uses when generating video with Vidu models on Scenario via MCP: text-to-video, image-to-video from a single still, start and end frame interpolation, reference-to-video keeping…

  • Works in 6 steps: search with target="models",… → model_schema_get with that id: duration… → upload_asset the start and end stills… → …
  • Generating video with Vidu models on Scenario via MCP: text-to-video
  • SKILL.md covers Overview, Quick reference, Tier picks the caps and Two prompt shapes, plus 2 more sections
  • Calls npx

What it does

Scenario Vidu is an agent skill from scenario-labs/skills. Use when generating video with Vidu models on Scenario via MCP: text-to-video, image-to-video from a single still, start and end frame interpolation, reference-to-video keeping characters and props consistent from up to 7 reference images or reference videos, anime-style motion, movement amplitude control, or deciding between Q3, Q2, Q1, and 2.0 tiers. Keywords: Vidu Q3 Pro, Q2 Turbo, Shengshu Technology, T2V, I2V, R2V, Reference2V, start end frame, background music toggle.

Its SKILL.md is about 1.5k tokens, which your agent loads only when the skill is triggered. It is a single SKILL.md file with no bundled scripts.

It sits in Media & Creative, covering AI video generation. It works with Model Context Protocol. The repository describes itself as: Get production-ready images, video, audio, and 3D from any AI agent: skills that pick the right model, price before spending, and keep characters and brands consistent through… The licence is MIT.

When your agent uses it

  • Generating video with Vidu models on Scenario via MCP: text-to-video
  • Image-to-video from a single still
  • Start and end frame interpolation
  • Reference-to-video keeping characters and props consistent from up to 7 reference images

Example prompts

  • “/scenario-vidu”

Requirements

  • Node.js

Workflow steps

6 steps, taken from the first numbered list in SKILL.md.

  1. search with target="models", query="vidu", public=true. Prefer the newest tier in the mode you need, e.g. model_vidu-i2v-q3-pro (a live…
  2. model_schema_get with that id: duration bounds, resolutions, defaults.
  3. upload_asset the start and end stills (see the scenario skill) for two asset ids.
  4. model_run with that model_id, dry_run=true, and parameters={"generationType": "start_end_to_video", "images": ["asset_start"…
  5. Re-run model_run with wait=false, then jobs_wait with the returned job id, re-called with pending_job_ids on timeout, never a second…
  6. asset_display the output and check both anchor frames landed.

What it can do on your machine

Read from SKILL.md and the folder at commit f6f8ab7. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Shell commands in SKILL.md call:

    • npx

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md. Its commands use npx, which can reach the network depending on how they are called.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Scenario Vidu loads about 1.5k tokens when it runs. Until then it costs about 123 tokens; SKILL.md has 785 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~123
When it runs · the whole SKILL.md, loaded when a task matches
~1.5k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from scenario-labs/skills at commit f6f8ab7, republished under its MIT licence (© scenario-labs). 785 words, ~1,456 tokens.

Download SKILL.mdSave it as .claude/skills/scenario-vidu/SKILL.md (or your agent's skills folder).
name
scenario-vidu
description
Use when generating video with Vidu models on Scenario via MCP: text-to-video, image-to-video from a single still, start and end frame interpolation, reference-to-video keeping characters and props consistent from up to 7 reference images or reference videos, anime-style motion, movement amplitude control, or deciding between Q3, Q2, Q1, and 2.0 tiers. Keywords: Vidu Q3 Pro, Q2 Turbo, Shengshu Technology, T2V, I2V, R2V, Reference2V, start end frame, background music toggle.
license
MIT

Scenario Vidu Video

Overview

Vidu, Shengshu Technology's video family on Scenario, ships one model per mode per tier instead of folding every mode into one model: the model name carries both the mode (T2V, I2V, Reference2V) and the tier (Q3, Q2, Q1, 2.0). Fifteen members were live at authoring time, so search the family, pick by name, and treat model_schema_get as the contract, since caps disagree at every tier boundary.

Connection and the core loop: see the scenario skill in this repo; model-agnostic video work: the scenario-video skill. If a sibling skill named here is missing from your available skills, ask the user to install it (npx skills add scenario-labs/skills --skill <name>); unattended, proceed from tool schemas and flag the gap.

Quick reference

Choose the model line by what you hold (input names from the live schemas):

You holdLineKey inputs
Text onlyT2Vprompt, aspectRatio, duration, resolution
One still to animateI2VgenerationType: "image_to_video", images with one id
Start and end framesI2VgenerationType: "start_end_to_video", images with two
Identity referencesReference2Vimages, up to 7

In start/end mode array order decides: the first id opens the clip, the second closes it, and generationType must match the image count. No I2V member exposes aspectRatio: geometry follows the source stills. Reference2V carries identity (characters, props, palette), not the opening state; to open on an exact frame, switch to an I2V member. One Reference2V member (Q2 Pro at authoring time) also took videos: up to 2 reference videos (one of 8 seconds or two of 5, 100MB each, aspect between 1:4 and 4:1), requiring images or videos.

Every member takes seed. audio (default false) is a background-music toggle, not sound design, and on most Q2 members it silently does nothing at 9 or 10 seconds.

Tier picks the caps

At authoring time: Q3 ran 1 to 16 seconds; Q2 ran 1 to 10, with start/end mode capped at 8; 2.0 I2V took exactly 4 or 8 seconds, and 8 only at 720p; Q1 exposed no duration or resolution at all, and 2.0 Reference2V no duration. movementAmplitude (auto, small, medium, large) exists only on the Q1 and 2.0 lines, and the general or anime style selector only on Q1 I2V members; on Q2 and Q3, motion energy and style go in the prompt. Turbo and Fast variants cut latency at the same duration and resolution caps as their Pro siblings. duration and resolution both move cost, several-fold on one member alone, so dry_run the exact job before a batch.

Two prompt shapes

T2V and I2V want a 120 to 180 word single-shot paragraph in present tense: scene and lighting, one or two motions, one camera move in plain verbs (pan, dolly, zoom in), emotion as physical cues ("her shoulders slump"), no cuts, and in I2V nothing that is not already in the still. Reference2V wants the opposite: one short sentence, roughly 20 to 40 words, naming which element comes from which reference by position, "the character from image1 wearing the armor from image2, walking through a snowstorm". Prompt length caps differ per member, 2000 to 5000 characters.

Show full SKILL.md (270 more words)Show less

Worked example: bridging two keyframes

  1. search with target="models", query="vidu", public=true. Prefer the newest tier in the mode you need, e.g. model_vidu-i2v-q3-pro (a live hit at authoring time: re-discover each session).
  2. model_schema_get with that id: duration bounds, resolutions, defaults.
  3. upload_asset the start and end stills (see the scenario skill) for two asset ids.
  4. model_run with that model_id, dry_run=true, and parameters={"generationType": "start_end_to_video", "images": ["asset_start", "asset_end"], "prompt": "The knight slowly lowers his sword as dusk settles over the courtyard. The camera dollies in on his face.", "duration": 8, "resolution": "1080p"}; expand the prompt to the full paragraph shape above in real runs, and re-estimate after any change to duration or resolution.
  5. Re-run model_run with wait=false, then jobs_wait with the returned job id, re-called with pending_job_ids on timeout, never a second model_run.
  6. asset_display the output and check both anchor frames landed.

Common mistakes

  • Two images with generationType: "image_to_video", or one with start/end: the type must match the image count, and order sets start versus end.
  • Expecting a Reference2V clip to open on a reference image: references fix identity, not frame one; use an I2V member for exact opening frames.
  • Carrying caps across tiers: 16 seconds is Q3 only; 2.0 I2V accepts only 4 or 8, and 8 forces 720p.
  • Prompting dialogue or sound effects: audio only adds background music, and it drops out at 9 or 10 seconds on most Q2 members.
  • A 150-word paragraph on Reference2V, or a bare sentence on T2V: the two modes want opposite prompt shapes.
  • Passing videos to any Reference2V member: at authoring time only one took reference videos; read the schema first.

© scenario-labs, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

Just SKILL.md in skills/scenario-vidu of scenario-labs/skills.

Open the folder on GitHubat commit f6f8ab7

Compare with similar skills

Scenario Vidu next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Scenario Vidu compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Scenario Vidu this skillscenario-labs/skills946—~1.5kAutomated safety check: PassMIT
Clipmivo VideoBarneyD66/clipmivo-tools142—~945Automated safety check: PassMIT
ComfyUI Local DriverSlavaSexton/ComfyUI-Agent-Kit105—~12kAutomated safety check: PassApache-2.0
Generate VideoArcReel/ArcReel5.4k—~1.3kAutomated safety check: WarnAGPL-3.0
Generate StoryboardArcReel/ArcReel5.4k—~816Automated safety check: PassAGPL-3.0
Manage ProjectArcReel/ArcReel5.4k—~1.6kAutomated safety check: PassAGPL-3.0

Similar skills

  • Clipmivo Video

    BarneyD66/clipmivo-tools

    Create and manage AI video tasks through ClipmivoAI using its MCP server, CLI or REST API.

    142 GitHub stars~945 tokensUpdated 25 days ago
    Media & CreativeAuto-check passed
  • ComfyUI Local Driver

    SlavaSexton/ComfyUI-Agent-Kit

    Drives a local ComfyUI install over its HTTP API to generate and edit images, video and audio, with per-model prompt recipes and workflow guidance.

    105 GitHub stars~12k tokensUpdated 1 mo ago
    Media & CreativeAuto-check passed
  • Generate Video

    ArcReel/ArcReel

    为分镜或自包含视频单元生成视频。当用户要求生成或重做视频时使用;支持整集、单项与批量自选. An agent skill from ArcReel/ArcReel.

    5.4k GitHub stars~1.3k tokensUpdated yesterday
    Media & CreativeAuto-check: warnings
  • Generate Storyboard

    ArcReel/ArcReel

    为分镜生成分镜图。当用户说"生成分镜"、"预览分镜画面"、想重新生成某些分镜图、或剧本中有分镜缺少分镜图时使用。自动保持角色和画面连续性。

    5.4k GitHub stars~816 tokensUpdated yesterday
    Media & CreativeAuto-check passed
  • Manage Project

    ArcReel/ArcReel

    项目管理工具集。使用场景:新增/修改角色/场景/道具到 project.json(经 patchproject 工具,按 table+name upsert)、级联重命名资产(renameasset 工具)、合并同一身份被重复登记的资产(mergeasset 工具)、写顶层 settings 字段、编辑项目概述…

    5.4k GitHub stars~1.6k tokensUpdated yesterday
    Media & CreativeAuto-check passed
  • Image Edit Workbench

    henjicc/Henji-AI

    在痕迹AI修图、调色、抠出或选中主体、移除物体、修补瑕疵、编辑图层与蒙版、导出图片或流转图片产物时使用。视频时间线与成片用 video-edit-workbench;写代码画面用 video-edit-code-creation。

    254 GitHub stars~502 tokensUpdated today
    Media & CreativeAuto-check passed

More from scenario-labs/skills

All 146 skills in this repo
  • Scenario Blender Grease Pencil

    scenario-labs/skills

    A skill your agent uses when drawing or animating with Grease Pencil in Blender 5.x from Python: 2D or 2.5D illustration, frame-by-frame animation, a cutout or part-based 2D character, strokes with…

    946 GitHub stars~4.5k tokensUpdated yesterday
    Auto-check passed
  • Scenario Blender Hair

    scenario-labs/skills

    A skill your agent uses when grooming hair or fur in Blender with hair curves, such as a character hairstyle, animal fur, procedural fur in geometry nodes, or hair cards and mesh hair for games.

    946 GitHub stars~4.7k tokensUpdated yesterday
    Auto-check passed
  • A skill your agent uses when lighting, rendering or compositing in Blender: light a character, product or hero shot, interior at dusk or night, three-point or motivated lighting, sun and sky, HDRI…

    946 GitHub stars~5k tokensUpdated yesterday
    Auto-check passed
  • Scenario Chatgpt Pet Create

    scenario-labs/skills

    A skill your agent uses when creating a ChatGPT pet or Codex pet with Scenario: hatching an animated companion from a text idea, a character, mascot or brand cue, or reference photos and art; making…

    946 GitHub stars~3.6k tokensUpdated yesterday
    Auto-check passed
  • Scenario Godot Animation

    scenario-labs/skills

    A skill your agent uses when animating characters or scenes in Godot 4.7: AnimationPlayer clips and RESET, AnimationTree state machines and blend spaces built in code, Mixamo or glTF import, loop…

    946 GitHub stars~4.7k tokensUpdated yesterday
    Auto-check passed
  • Scenario Godot Audio

    scenario-labs/skills

    A skill your agent uses when adding or fixing sound in Godot 4.7: audio buses and effects, volume sliders, 'too many sounds', combat audio with hundreds of enemies, sounds clipping or distorting, 3D…

    946 GitHub stars~4.7k tokensUpdated yesterday
    Auto-check passed

Questions about Scenario Vidu

What does Scenario Vidu do?

A skill your agent uses when generating video with Vidu models on Scenario via MCP: text-to-video, image-to-video from a single still, start and end frame interpolation, reference-to-video keeping…. Scenario Vidu is an agent skill from scenario-labs/skills.0 tiers.

When should I use Scenario Vidu?

Scenario Vidu fits situations like: generating video with Vidu models on Scenario via MCP: text-to-video; image-to-video from a single still; start and end frame interpolation; reference-to-video keeping characters and props consistent from up to 7 reference images.

How do I install Scenario Vidu in Claude Code?

Run `npx skills add scenario-labs/skills --skill scenario-vidu -a claude-code`. Or copy the skill folder (skills/scenario-vidu in scenario-labs/skills) into .claude/skills/scenario-vidu in your project. Claude Code loads it when a task matches its description.

How do I install Scenario Vidu in Codex?

Run `npx skills add scenario-labs/skills --skill scenario-vidu -a codex`. Or copy the skill folder (skills/scenario-vidu in scenario-labs/skills) into .agents/skills/scenario-vidu in your project. Codex loads it when a task matches its description.

Can I use Scenario Vidu in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add scenario-labs/skills --skill scenario-vidu -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/scenario-vidu, .gemini/skills/scenario-vidu, .github/skills/scenario-vidu and .opencode/skills/scenario-vidu in your project.

What does Scenario Vidu need to run?

Going by SKILL.md and its folder, Scenario Vidu needs the command-line tools its instructions call (npx). Our summary lists: Node.js.

Does Scenario Vidu access the network?

SKILL.md contains no URLs. Its commands use npx, which can reach the network depending on how they are called. This is read from the text; nothing was executed.

Is Scenario Vidu safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Scenario Vidu use?

Scenario Vidu is published under the MIT licence (declared in SKILL.md). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Scenario Vidu use?

About 1.5k tokens (SKILL.md is roughly 5.8k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Scenario Vidu?

Skills that share tags, products or a category with Scenario Vidu: Clipmivo Video (BarneyD66/clipmivo-tools, 142 stars), ComfyUI Local Driver (SlavaSexton/ComfyUI-Agent-Kit, 105 stars), Generate Video (ArcReel/ArcReel, 5.4k stars) and Generate Storyboard (ArcReel/ArcReel, 5.4k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Scenario Vidu?

scenario-labs (a GitHub organization) maintains it in scenario-labs/skills, which has 946 GitHub stars. The repository holds 146 skills in this directory. The repository was last updated on October 10, 2026.

Source: scenario-labs/skills on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.