Agent skill

Direct Video Creation

by openvetta in openvetta/open-vetta

Turn creative intent into a production-ready AI video brief, treatment, script/beat plan, shot cards, animatic/keyframe plan, node workflow, and model-profile prompts.

Apache-2.0Auto-check passedMedia & Creative

Install Direct Video Creation

skills CLI
$ npx skills add openvetta/open-vetta --skill direct-video-creation -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install openvetta/open-vetta direct-video-creation --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/openvetta/open-vetta.git skills-src && mkdir -p .claude/skills && cp -r skills-src/packages/plugins/presets/content-creation/agent/skills/direct-video-creation .claude/skills/direct-video-creation && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
direct-video-creation
GitHub stars
290
Token cost
~1.9k tokens
SKILL.md length
584 words
Files
31 (incl. references)
Skills in repo
21
Repo updated
First seen
Licence
Apache-2.0

At a glance

Turn creative intent into a production-ready AI video brief, treatment, script/beat plan, shot cards, animatic/keyframe plan, node workflow, and model-profile prompts.

  • Works in 6 steps: Inspect the project and capabilities. → If no concept, story, or scene exists,… → Classify the request as… → …
  • Dialogue scenes
  • SKILL.md covers Route the task, Brief requirements and Choose a generation mode
  • Instructions only: no scripts, shell commands, URLs or credentials in SKILL.md

What it does

Direct Video Creation is an agent skill from openvetta/open-vetta. Turn creative intent into a production-ready AI video brief, treatment, script/beat plan, shot cards, animatic/keyframe plan, node workflow, and model-profile prompts. Use for text-to-video, image-to-video, multi-shot or dialogue scenes, reference-video transformation, video editing or continuation, product/logo/jewelry/fashion/social/UGC films, fight scenes, character stories, cooking tutorials, music videos, award or freeze effects, talking-character clips, drone or one-shot footage, long-form-to-short clipping…

Its SKILL.md is about 1.9k tokens, which your agent loads only when the skill is triggered. The skill folder holds 32 other files, including reference files (for example `agents/openai.yaml`, `references/animate-still-method.md` and `references/animatic-keyframes.md`).

It sits in Media & Creative, covering AI video generation. The repository describes itself as: Open-source, local-first AI agent for coding and real work. BYOK models, MCP, skills, plugins, workflows, and private knowledge bases. The licence is Apache-2.0.

When your agent uses it

  • Dialogue scenes
  • Reference-video transformation
  • Product/logo/jewelry/fashion/social/UGC films
  • Character stories

Example prompts

  • “/direct-video-creation”

Workflow steps

6 steps, taken from the first numbered list in SKILL.md.

  1. Inspect the project and capabilities.
  2. If no concept, story, or scene exists, invoke $develop-creative-concept first.
  3. Classify the request as treatment/script, storyboard/animatic, text-to-video, image-to-video, reference transformation, edit/continuation…
  4. Build the brief and shot plan before adding generation nodes.
  5. Generate a low-cost proof shot when direction is uncertain; expand only after the direction is accepted.
  6. Review the actual output with $review-content-quality before recommending a retry.

What it can do on your machine

Read from SKILL.md and the folder at commit b983179. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    No scripts in the folder and no shell commands in SKILL.md.

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    Links to these hosts (documentation or services it may open):

    • github.com

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Direct Video Creation loads about 1.9k tokens when it runs, and up to ~24k if it reads all its reference files. Until then it costs about 168 tokens; SKILL.md has 584 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~168
When it runs · the whole SKILL.md, loaded when a task matches
~1.9k
With references · SKILL.md plus every file in references/, read only if the agent opens them
~24k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from openvetta/open-vetta at commit b983179, republished under its Apache-2.0 licence (© openvetta). 584 words, ~1,929 tokens.

Download SKILL.mdSave it as .claude/skills/direct-video-creation/SKILL.md (or your agent's skills folder). This skill also uses 30 other files; get the full folder from GitHub.
name
direct-video-creation
description
Turn creative intent into a production-ready AI video brief, treatment, script/beat plan, shot cards, animatic/keyframe plan, node workflow, and model-profile prompts. Use for text-to-video, image-to-video, multi-shot or dialogue scenes, reference-video transformation, video editing or continuation, product/logo/jewelry/fashion/social/UGC films, fight scenes, character stories, cooking tutorials, music videos, award or freeze effects, talking-character clips, drone or one-shot footage, long-form-to-short clipping plans, storyboards, prompt audits, camera/light/sound direction, race/chase/kinetic montage, or improving a video-generation node.

Direct AI video creation

Use $operate-content-workflow for all inspection and mutations. This skill supplies creative decisions; runtime capabilities remain the source of truth for what can be executed.

Route the task

  1. Inspect the project and capabilities.
  2. If no concept, story, or scene exists, invoke $develop-creative-concept first.
  3. Classify the request as treatment/script, storyboard/animatic, text-to-video, image-to-video, reference transformation, edit/continuation, single shot, or multi-shot sequence.
  4. Build the brief and shot plan before adding generation nodes.
  5. Generate a low-cost proof shot when direction is uncertain; expand only after the direction is accepted.
  6. Review the actual output with $review-content-quality before recommending a retry.

Read only the references needed for the task:

Show full SKILL.md (265 more words)Show less

Brief requirements

Record purpose, audience, publishing surface, duration, ratio, subject, environment, style, continuity anchors, action, camera, pacing, audio intent, references, constraints, deliverables, and acceptance criteria. Infer reversible defaults; ask only when a missing choice changes cost, supplied references, or the core deliverable.

Keep each generator focused on one visually coherent shot unless the inspected mode explicitly supports multiple timestamped stages or shots in one generation. In that mode, put consecutive time windows inside the video prompt. Otherwise create independent shot nodes and return their intended order. Name the changed variable in every intentional variation.

Choose a generation mode

  • Select the method from the authorities the shot must preserve, using video-strategy-selection.md before considering model availability.
  • Use the strategy-specific prompt plan kind required by that method, even when strategy="automatic".
  • Use first/last-frame behavior only when the inspected mode exposes distinct frame roles and both static endpoint plans are feasible.
  • Use omni-reference or video transformation only when the inspected input slots accept the required media kinds.
  • Prefer a short low-cost validation shot before a larger set when creative direction is uncertain.

When materializing the plan on the canvas, use $operate-content-workflow's high-level configure_video_shot contract. Pass concrete asset IDs for asset collections, reference upstream image/video generators by future output, and let capability resolution compile business roles to provider slots. Use low-level configure_generation only to preserve or repair an existing technical role configuration. Never express first frame, last frame, visual reference, motion reference, or source video as an unlabelled edge.

This method is an original Vetta adaptation informed by Generative-Media-Skills (MIT), visual-skills by Serge Shima (CC BY 4.0, https://github.com/smixs/visual-skills), and ViMax (MIT).

© openvetta, Apache-2.0. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

SKILL.md and 30 other files (references) in packages/plugins/presets/content-creation/agent/skills/direct-video-creation of openvetta/open-vetta.

  • SKILL.md
  • agents/openai.yaml
  • references/animate-still-method.md
  • references/animatic-keyframes.md
  • references/camera-light-sound-vocabulary.md
  • references/camera-social-and-clipping-video-recipes.md
  • references/character-performance-and-ugc-video-recipes.md
  • references/continuity-and-references.md
  • references/dramaturgy-and-shot-design.md
  • references/failure-repairs.md
  • references/first-last-frame-method.md
  • references/generation-timeline-and-storyboard.md
  • references/generation-timeline-examples.md
  • references/genre-and-montage-patterns.md
  • references/kinetic-speed.md
  • references/model-prompt-profiles.md
  • references/narrative-action-and-tutorial-video-recipes.md
  • references/omni-reference-method.md
  • references/product-brand-and-logo-video-recipes.md
  • … and 12 more

Open the folder on GitHubat commit b983179

Compare with similar skills

Direct Video Creation next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Direct Video Creation compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Direct Video Creation this skillopenvetta/open-vetta290—~1.9kAutomated safety check: PassApache-2.0
Video Generationbytedance/deer-flow84k3 repos~1.4kAutomated safety check: PassMIT
Video Cover Imageitwanger/toBeBetterJavaer18k—~3.3kAutomated safety check: PassNone
Seedancesongguoxs/seedance-prompt-skill2.9k1 repos~2.5kAutomated safety check: PassNone
HyperFrames Video Entry Pointheygen-com/hyperframes60k3 repos~5.2kAutomated safety check: PassApache-2.0
Lanshu Create AI Presenter Videocclank/lanshu-create-ai-presenter-video2.6k—~3.6kAutomated safety check: PassMIT

Similar skills

  • Video Generation

    bytedance/deer-flow

    Generates short videos from a structured JSON prompt, optionally guided by a reference image used as the first or last frame.

    84k GitHub starsUsed in 3 repos~1.4k tokens
    Media & CreativeAuto-check passed
  • Video Cover Image

    itwanger/toBeBetterJavaer

    Generate matched 3:4, 16:9, and 4:3 short-video cover images from toBeBetterJavaer video scripts or AI/Java technical topics.

    18k GitHub stars~3.3k tokensUpdated today
    Media & CreativeAuto-check passed
  • Seedance

    songguoxs/seedance-prompt-skill

    This skill should be used when the user asks to "generate video prompts", "create Seedance prompts", "write video descriptions", mentions "Seedance", "seedance", "即梦", "即梦平台", "视频提示词", "视频生成"…

    2.9k GitHub starsUsed in 1 repo~2.5k tokens
    Media & CreativeAuto-check passed
  • HyperFrames Video Entry Point

    heygen-com/hyperframes

    Entry point for making, editing and rendering videos from HTML compositions with HyperFrames, routing each request to the right workflow.

    60k GitHub starsUsed in 3 repos~5.2k tokens
    Media & CreativeAuto-check passed
  • Lanshu Create AI Presenter Video

    cclank/lanshu-create-ai-presenter-video

    Turn a topic or finished script into a complete, publish-ready explainer video — led by an AI presenter from an authorized adult presenter image, or performed in one of nine visual explainer styles…

    2.6k GitHub stars~3.6k tokensUpdated 4 days ago
    Media & CreativeAuto-check passed
  • Video Shots

    eternityspring/reelbench-skills

    拉片:把一条成片拆成逐镜头的分析表——每个镜头的时长、景别、类别、运镜、画面. An agent skill from eternityspring/reelbench-skills.

    878 GitHub starsUsed in 1 repo~1.8k tokens
    Media & CreativeAuto-check: notes

More from openvetta/open-vetta

All 21 skills in this repo
  • Mobile Android Design

    openvetta/open-vetta

    Master Material Design 3 and Jetpack Compose patterns for building native Android apps.

    290 GitHub starsUsed in 3 repos~950 tokens
    Auto-check passed
  • Publish Ability

    openvetta/open-vetta

    Publish a skill, scene, MCP server, plugin, or bundle to the Vetta ability marketplace.

    290 GitHub stars~1.6k tokensUpdated yesterday
    Auto-check passed
  • Plugin Workbench

    openvetta/open-vetta

    Create, implement, build, pack, install, reload, and manage Vetta desktop plugins for non-developers.

    290 GitHub stars~1.9k tokensUpdated yesterday
    Auto-check passed
  • Create Skill

    openvetta/open-vetta

    Create or update Vetta-compatible Agent Skills. An agent skill from openvetta/open-vetta.

    290 GitHub stars~1.6k tokensUpdated yesterday
    Auto-check passed
  • Mobile UI Design

    openvetta/open-vetta

    Design mobile-style HTML pages (iOS / Android) for preview in the Mobile UI Preview panel.

    290 GitHub stars~1.6k tokensUpdated yesterday
    Auto-check passed
  • Remotion Video

    openvetta/open-vetta

    Create, modify, and render Remotion React video projects in the current Vetta conversation workspace.

    290 GitHub stars~744 tokensUpdated yesterday
    Auto-check passed

Questions about Direct Video Creation

What does Direct Video Creation do?

Turn creative intent into a production-ready AI video brief, treatment, script/beat plan, shot cards, animatic/keyframe plan, node workflow, and model-profile prompts. Direct Video Creation is an agent skill from openvetta/open-vetta. Turn creative intent into a production-ready AI video brief, treatment, script/beat plan, shot cards, animatic/keyframe plan, node workflow, and model-profile prompts.

When should I use Direct Video Creation?

Direct Video Creation fits situations like: dialogue scenes; reference-video transformation; product/logo/jewelry/fashion/social/UGC films; character stories.

How do I install Direct Video Creation in Claude Code?

Run `npx skills add openvetta/open-vetta --skill direct-video-creation -a claude-code`. Or copy the skill folder (packages/plugins/presets/content-creation/agent/skills/direct-video-creation in openvetta/open-vetta) into .claude/skills/direct-video-creation in your project. Claude Code loads it when a task matches its description.

How do I install Direct Video Creation in Codex?

Run `npx skills add openvetta/open-vetta --skill direct-video-creation -a codex`. Or copy the skill folder (packages/plugins/presets/content-creation/agent/skills/direct-video-creation in openvetta/open-vetta) into .agents/skills/direct-video-creation in your project. Codex loads it when a task matches its description.

Can I use Direct Video Creation in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add openvetta/open-vetta --skill direct-video-creation -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/direct-video-creation, .gemini/skills/direct-video-creation, .github/skills/direct-video-creation and .opencode/skills/direct-video-creation in your project.

What does Direct Video Creation need to run?

SKILL.md names no scripts, command-line tools or credentials: Direct Video Creation is instructions for the agent only.

Does Direct Video Creation access the network?

SKILL.md names 1 domain. As links in the text: github.com. This is read from the text; nothing was executed.

Is Direct Video Creation safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Direct Video Creation use?

Direct Video Creation is published under the Apache-2.0 licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Direct Video Creation use?

About 1.9k tokens (SKILL.md is roughly 7.7k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full. Its references folder adds about 22k tokens, read only when the agent opens those files.

What are the alternatives to Direct Video Creation?

Skills that share tags, products or a category with Direct Video Creation: Video Generation (bytedance/deer-flow, 84k stars), Video Cover Image (itwanger/toBeBetterJavaer, 18k stars), Seedance (songguoxs/seedance-prompt-skill, 2.9k stars) and HyperFrames Video Entry Point (heygen-com/hyperframes, 60k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Direct Video Creation?

openvetta (a GitHub organization) maintains it in openvetta/open-vetta, which has 290 GitHub stars. The repository holds 21 skills in this directory. The repository was last updated on October 9, 2026.

Source: openvetta/open-vetta on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.