Agent skill

Video

by smixs in smixs/visual-skills

A skill your agent uses whenever the user asks to create, improve, audit, or split prompts for AI video generators (Seedance, Kling, Veo, Runway, Luma, Pika, Sora, any image-to-video system).

CC-BY-4.0Auto-check passedMedia & Creative

Install Video

skills CLI
$ npx skills add smixs/visual-skills --skill video -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install smixs/visual-skills video --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/smixs/visual-skills.git skills-src && mkdir -p .claude/skills && cp -r skills-src/video .claude/skills/video && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
video
GitHub stars
490
Token cost
~2.5k tokens
SKILL.md length
1,128 words
Files
13 (incl. references)
Skills in repo
2
Repo updated
First seen
Licence
CC-BY-4.0

At a glance

A skill your agent uses whenever the user asks to create, improve, audit, or split prompts for AI video generators (Seedance, Kling, Veo, Runway, Luma, Pika, Sora, any image-to-video system).

  • Works in 5 steps: always read first → dramaturgy.md → always read second → universal-rules.md → pick the model and read one model file → …
  • The user asks to create
  • Instructions only: no scripts, shell commands, URLs or credentials in SKILL.md
  • Split prompts for AI video generators (Seedance

What it does

Video is an agent skill from smixs/visual-skills. Use this skill whenever the user asks to create, improve, audit, or split prompts for AI video generators (Seedance, Kling, Veo, Runway, Luma, Pika, Sora, any image-to-video system). The skill also covers storyboards, shot lists, director treatments, dynamic montage, multi-clip story structure, camera direction, lighting, blocking, pacing, character continuity, dialogue, and sound design. Trigger even when the user says things like "придумай сцену для видео", "разбей на склейки", "сделай раскадровку", "улучши…

Its SKILL.md is about 2.5k tokens, which your agent loads only when the skill is triggered. The skill folder holds 13 other files, including reference files (for example `references/animatic-keyframes.md`, `references/camera-lighting-vocabulary.md` and `references/dramaturgy.md`).

It sits in Media & Creative, covering AI video generation. It works with Seedance, xAI Grok, Claude Agent SDK and Google Gemini. The repository describes itself as: AI film director skills for agents: cinematic dramaturgy (Murch, blocking, montage) + exact prompt syntax for Seedance 2.5, Kling 3.0 Turbo/Omni, Veo 3.1, Nano Banana 2, GPT… The licence is CC-BY-4.0.

When your agent uses it

  • The user asks to create
  • Split prompts for AI video generators (Seedance
  • Any image-to-video system)
  • Even when the user says things like придумай сцену для видео

Example prompts

  • “улучши промпт для Kling”
  • “как снять X в AI-видео”
  • “prompt”
  • “/video”

Workflow steps

5 steps, taken from the step headings in SKILL.md.

  1. always read first → dramaturgy.md
  2. always read second → universal-rules.md
  3. pick the model and read one model file
  4. task-shaped reading (load only those that match)
  5. apply the dramaturgy check and the three-detail check

What it can do on your machine

Read from SKILL.md and the folder at commit 92be33a. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    No scripts in the folder and no shell commands in SKILL.md.

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    Links to these hosts (documentation or services it may open):

    • github.com
    • t.me
    • sergeshima.com
    • aimasters.me

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Video loads about 2.5k tokens when it runs, and up to ~52k if it reads all its reference files. Until then it costs about 219 tokens; SKILL.md has 1,128 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~219
When it runs · the whole SKILL.md, loaded when a task matches
~2.5k
With references · SKILL.md plus every file in references/, read only if the agent opens them
~52k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from smixs/visual-skills at commit 92be33a, republished under its CC-BY-4.0 licence (© smixs). 1,128 words, ~2,463 tokens.

Download SKILL.mdSave it as .claude/skills/video/SKILL.md (or your agent's skills folder). This skill also uses 12 other files; get the full folder from GitHub.
name
video
description
Use this skill whenever the user asks to create, improve, audit, or split prompts for AI video generators (Seedance, Kling, Veo, Runway, Luma, Pika, Sora, any image-to-video system). The skill also covers storyboards, shot lists, director treatments, dynamic montage, multi-clip story structure, camera direction, lighting, blocking, pacing, character continuity, dialogue, and sound design. Trigger even when the user says things like "придумай сцену для видео", "разбей на склейки", "сделай раскадровку", "улучши промпт для Kling", "переведи сценарий в промпты", "как снять X в AI-видео", or shares a prompt and asks to fix it — even if they never say the word "prompt". Do NOT use for still images (use the sibling image skill), for editing, colour-grading or assembling footage that already exists, or for writing feature-length screenplays with no generation step.
license
CC-BY-4.0 (attribution required — Serge Shima, github.com/smixs/visual-skills)

AI Director, Screenwriter & Editor

Hybrid role. You direct (see frame, emotion, motivated camera), write (build beat, action, consequence, final image), and edit (cut rhythm, protect continuity, drive montage). Prompt engineering is fourth — it serves the first three.

A beautiful frame without dramaturgy is wallpaper. A dramaturgically clean prompt without details is mush. The whole craft of this skill lives in the reference files. The body of this SKILL.md is intentionally thin so you cannot fake a result by reading it alone.

Route first — is this actually a video-prompt task?

  • No idea or script yet (user wants a concept, a Big Idea, a campaign, an ad scenario — not a prompt): if the creative-director skill is installed, start there — it develops ideas and scripts for commercials and beyond (github.com/smixs/creative-director-skill). Come back here once there is a script to shoot.
  • Still keyframes, character sheets, animatic panels to feed the video pipeline: use the sibling image skill, then return with the keyframes.
  • A script or scene exists and needs prompts — this skill. Continue below.

Mandatory reading order — DO NOT WRITE A PROMPT WITHOUT THIS

Past attempts to write prompts directly from this skill body produced lazy, mush-prone results. The fix is structural: the process lives only in the reference files, and you load them in this order before producing output. Skipping a step silently degrades the result — the model cannot tell that a shot is wallpaper, only the writer can, and only by applying the rules from these files.

For every video prompt request, load the files in this order:

Step 1 — always read first → dramaturgy.md

Scene formula. Details Law (the second core law, most violated). Murch Rule of Six. Three-jobs rule. Five anchors. Blocking, staging, environment as pressure. Three-layer storyboard. 14-field shot card. Rhythm ladder. Dramaturgy check.

You cannot decide whether a prompt is ready without running the dramaturgy check from this file.

Step 2 — always read second → universal-rules.md

U1–U12 universal rules that apply to every video model: prompt skeleton, weight-at-start, show-don't-tell, lens language, character anchor, contradictions, duration discipline, final image rule, three-detail check.

Step 3 — pick the model and read one model file

Use this short selector. The full reasoning is in the chosen file.

Cue from the user / taskRead
Seedance, ByteDance, Doubao, Jimeng, multi-shot in one clip, --resolution, --duration, --camerafixed, "Cut to", @img1, fast multi-shot dramaseedance.md
Seedance 2.5 production work: 30s single-pass, 50-slot reference kits, video editing / partial re-render, extension, Ultra Long (30-180s), 3D blockout / green screen, @Image N, { } dialogue markersseedance.md + seedance-25.md
Kling, Kuaishou, Element Binding, Motion Brush, Motion Control, dedicated negative prompt field, Kling 3.0 multi-shot with [Character A: ...] labels, native dialogue + lip-sync, 15s, Turbo (cheap lip-sync), Omni (references + editing, 4K)kling.md
Veo, Google video, dialogue / lip-sync, JSON prompts, synchronized SFX, commercial polish with voiceoverveo.md

Default if nothing in the request hints at a model:

  • Multi-shot narrative or fast montage drama → Seedance, or Kling 3.0 if dialogue is involved.
  • Dialogue / commercial polish / synchronized SFX → Veo, or Kling 3.0 for multi-character dialogue scenes up to 15s.
  • Character consistency across many social clips → Kling 2.6 Pro (cheaper) or Kling 3.0 (with in-prompt [Character A: ...] labels).
  • 10-15s continuous narrative with audio → Kling 3.0.
  • 15-30s continuous single-generation arc, heavy reference kits (up to 50 assets), editing or extending existing footage, 30-180s long-form → Seedance 2.5.
  • Face-heavy drama → Seedance 2.5 (realistic humans + lip-sync are its headline feature), Kling, or Veo. On a 2.0-only pipeline route faces to 1.5 Pro (2.0 filters human faces aggressively).

For a more detailed comparison (max clip length, audio support, character lock methods, motion brush, etc.), read the model file you picked. Do not load all three.

Show full SKILL.md (529 more words)Show less
Step 4 — task-shaped reading (load only those that match)
  • Storyboard / shot list / director treatment / "разбей на склейки" → role-modes.md. Determines whether you operate as Director, Screenwriter, or Editor for this turn.
  • Storyboard keyframes / опорные кадры / аниматик / animatic / still panels / key visuals to pitch a sequence → animatic-keyframes.md. The general method for turning a beat sheet into still panels (and then image-gen prompts) that read as story, drama and emotion without motion or faces.
  • Race / drift / drag / chase / speed / dynamic / kinetic montage, "гонщик", "раскадровка гонки", authentic-speed spot → race-and-speed.md. Specializes animatic-keyframes.md for the race domain — read that file first.
  • Commercial, music video, drama, action, fashion, UGC, product film, escalation / anxiety / discovery / catastrophe / product-drama montage → patterns-and-genres.md.
  • The shape of the story is not settled yet (no opponent, one single culminating moment, a deliberately open ending, or the brief reads as a set of nice frames) → patterns-and-genres.md §4, choosing the story arc. Read it before the beat map: laying a causal beat map on a story that has no causal spine is what makes a spot feel busy but empty.
  • Multi-clip continuity, fixing a broken prompt, known failure modes (one-take, face drift, melted hands, dialogue too fast) → fixes-and-skeletons.md.
  • Need precise framing / lens / movement / light / sound terms → camera-lighting-vocabulary.md.

If none match — proceed with steps 1-3 only.

Step 5 — apply the dramaturgy check and the three-detail check

Before returning anything, run both checks:

  • Dramaturgy check (dramaturgy.md §15): scene formula complete, three-detail check on every shot, three-jobs rule on every shot, motivated camera, readable geometry, five anchors named.
  • Three-detail audit (universal-rules.md §13): each shot owns environmental pressure + physical micro-action + sound or visual motif.

If any shot fails, fix before sending. This is the step the user has had to enforce repeatedly. Do not skip it.


Output

Choose the format the request actually asks for. Default to A if unclear.

  • A. Single prompt. One ready-to-copy prompt for one generation. Lead with model name + parameters in a short header.
  • B. Multi-clip prompts. Sequence of self-contained prompts, each repeating the full identity / style / continuity block (see universal-rules.md U7).
  • C. Storyboard. Table — Time, Shot, Function, Action, Camera, Light, Sound, Emotion. Every row is a 14-field shot card from dramaturgy.md §11, compressed.
  • D. Prompt audit. Given a user prompt, return: What works, What breaks generation, Missing direction, Continuity risks, Model-specific mismatches, Stronger version (rewritten prompt).
  • E. Director treatment. Core idea, Emotional arc, Visual motif, Rhythm, Camera language, Lighting, Sound, Ending image. (Treatment ≠ prompt.)
  • F. JSON (Veo only). Structured scene-by-scene continuity. See veo.md.

Default output language follows the user. The final AI prompt itself goes in English unless the user asks otherwise — Seedance, Kling, and Veo all perform better in English.


Final response style

Prefer: ready-to-copy prompts, clear section labels, production language, motivated camera and light direction, strict continuity blocks, model-specific syntax, direct fixes.

Avoid: long theory unless asked, academic lectures, vague inspiration, decorative jargon, "cinematic masterpiece" filler, prompts without camera and light, prompts without continuity, stacking more than two director references, abstract emotions without physical translation.

When in doubt about a model-specific detail — re-read the model file before writing the final prompt. It costs nothing and prevents bad output.


Author: Serge Shima (t.me/aimastersme · sergeshima.com · aimasters.me) · License: CC BY 4.0 — attribution required · Source: smixs/visual-skills

© smixs, CC-BY-4.0. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

SKILL.md and 12 other files (references) in video of smixs/visual-skills.

  • SKILL.md
  • references/animatic-keyframes.md
  • references/camera-lighting-vocabulary.md
  • references/dramaturgy.md
  • references/fixes-and-skeletons.md
  • references/kling.md
  • references/patterns-and-genres.md
  • references/race-and-speed.md
  • references/role-modes.md
  • references/seedance-25.md
  • references/seedance.md
  • references/universal-rules.md
  • references/veo.md

Open the folder on GitHubat commit 92be33a

Compare with similar skills

Video next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Video compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Video this skillsmixs/visual-skills490—~2.5kAutomated safety check: PassCC-BY-4.0
Seedance Storyboard Generatorliangdabiao/Seedance2-Storyboard-Generator2.5k—~2.2kAutomated safety check: PassNone
HiggsfieldOSideMedia/higgsfield-ai-prompt-skill707—~9.1kAutomated safety check: PassMIT
Nbcraftjieyefriic/nbcraft155—~2.8kAutomated safety check: PassMIT
AI Video Gencalesthio/OpenMontage66k—~3kAutomated safety check: PassAGPL-3.0
AI Media GeneratorHao0321/ai-media-generator258—~5.6kAutomated safety check: PassMIT

Similar skills

  • Seedance Storyboard Generator

    liangdabiao/Seedance2-Storyboard-Generator

    专业的Seedance 2.0平台AI视频脚本和分镜生成器。当用户要求:(1) 将文章/故事转换为视频脚本,(2) 生成Seedance 2.0分镜提示词,(3) 规划多集AI视频系列,(4) 为GPT-Image-2、Seedream、Nano Banana…

    2.5k GitHub stars~2.2k tokensUpdated 18 days ago
    Media & CreativeAuto-check passed
  • Higgsfield

    OSideMedia/higgsfield-ai-prompt-skill

    A skill your agent uses whenever the user asks anything about Higgsfield AI — writing or refining video/image prompts, choosing a model (Kling, Veo, Wan, Seedance, Minimax Hailuo, DoP, Soul, Nano…

    707 GitHub stars~9.1k tokensUpdated 12 days ago
    Media & CreativeAuto-check passed
  • Nbcraft

    jieyefriic/nbcraft

    Multi-backend Image + Video Generation CLI (nb command). An agent skill from jieyefriic/nbcraft.

    155 GitHub stars~2.8k tokensUpdated 5 mo ago
    Media & CreativeAuto-check passed
  • AI Video Gen

    calesthio/OpenMontage

    Generate AI videos from text prompts using multiple provider gateways.

    66k GitHub stars~3k tokensUpdated 5 days ago
    Media & CreativeAuto-check passed
  • AI Media Generator

    Hao0321/ai-media-generator

    為使用者產生高品質的 AI 生圖、生影片、生音樂提示詞,並在需要時透過瀏覽器自動化實際送到目標平台。涵蓋 OiiOii、Kling 3.0/O-series、Seedance 2.0/2.5、Suno v5.5、Seedream 5.0/4.0、Vidu Q3、Midjourney V8.1、Flux 1.1 Pro / Kontext、Runway Gen-4.5 /…

    258 GitHub stars~5.6k tokensUpdated 2 mo ago
    Media & CreativeAuto-check passed
  • AI Video Gen

    calesthio/OpenMontage

    Generate AI videos from text prompts using multiple provider gateways.

    66k GitHub stars~2.8k tokensUpdated 5 days ago
    Media & CreativeAuto-check passed

More from smixs/visual-skills

  • Image

    smixs/visual-skills

    Image prompting skill for Nano Banana (NBP/NB2) and GPT Image 2.5 (Flare/Sunburst).

    490 GitHub stars~2.1k tokensUpdated 23 days ago
    Auto-check passed

Questions about Video

What does Video do?

A skill your agent uses whenever the user asks to create, improve, audit, or split prompts for AI video generators (Seedance, Kling, Veo, Runway, Luma, Pika, Sora, any image-to-video system). Video is an agent skill from smixs/visual-skills. Use this skill whenever the user asks to create, improve, audit, or split prompts for AI video generators (Seedance, Kling, Veo, Runway, Luma, Pika, Sora, any image-to-video system).

When should I use Video?

Video fits situations like: the user asks to create; split prompts for AI video generators (Seedance; any image-to-video system); even when the user says things like придумай сцену для видео.

How do I install Video in Claude Code?

Run `npx skills add smixs/visual-skills --skill video -a claude-code`. Or copy the skill folder (video in smixs/visual-skills) into .claude/skills/video in your project. Claude Code loads it when a task matches its description.

How do I install Video in Codex?

Run `npx skills add smixs/visual-skills --skill video -a codex`. Or copy the skill folder (video in smixs/visual-skills) into .agents/skills/video in your project. Codex loads it when a task matches its description.

Can I use Video in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add smixs/visual-skills --skill video -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/video, .gemini/skills/video, .github/skills/video and .opencode/skills/video in your project.

What does Video need to run?

SKILL.md names no scripts, command-line tools or credentials: Video is instructions for the agent only.

Does Video access the network?

SKILL.md names 4 domains. As links in the text: github.com, t.me, sergeshima.com and aimasters.me. This is read from the text; nothing was executed.

Is Video safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Video use?

Video is published under the CC-BY-4.0 licence (declared in SKILL.md). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Video use?

About 2.5k tokens (SKILL.md is roughly 9.9k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full. Its references folder adds about 49k tokens, read only when the agent opens those files.

What are the alternatives to Video?

Skills that share tags, products or a category with Video: Seedance Storyboard Generator (liangdabiao/Seedance2-Storyboard-Generator, 2.5k stars), Higgsfield (OSideMedia/higgsfield-ai-prompt-skill, 707 stars), Nbcraft (jieyefriic/nbcraft, 155 stars) and AI Video Gen (calesthio/OpenMontage, 66k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Video?

smixs (a GitHub user) maintains it in smixs/visual-skills, which has 490 GitHub stars. The repository holds 2 skills in this directory. The repository was last updated on September 16, 2026.

Source: smixs/visual-skills on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.