Agent skill

H3 Prompt Writing

by vllm-project in vllm-project/vllm-omni

Write MiniMax H3 video generation prompts for T2VA, I2VA, FL2VA, L2VA, and Ref2VA.

Apache-2.0Auto-check passedMedia & Creative

Install H3 Prompt Writing

skills CLI
$ npx skills add vllm-project/vllm-omni --skill h3-prompt-writing -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install vllm-project/vllm-omni h3-prompt-writing --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/vllm-project/vllm-omni.git skills-src && mkdir -p .claude/skills && cp -r skills-src/.agents/skills/h3-prompt-writing .claude/skills/h3-prompt-writing && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
h3-prompt-writing
GitHub stars
7.1k
Used in
6 other repos
Token cost
~744 tokens
SKILL.md length
300 words
Files
5 (incl. references)
Skills in repo
20
Repo updated
First seen
Licence
Apache-2.0

At a glance

Write MiniMax H3 video generation prompts for T2VA, I2VA, FL2VA, L2VA, and Ref2VA.

  • Works in 5 steps: Read portable-workflow.md for scope,… → Identify the input mode: T2VA, I2VA,… → For base text/keyframe modes, read… → …
  • Rewriting multimodal requests into H3 prompt structures
  • SKILL.md covers Workflow, Base Modes, Full-Reference Mode and Output Rules, plus 1 more section
  • Instructions only: no scripts, shell commands, URLs or credentials in SKILL.md

What it does

H3 Prompt Writing is an agent skill from vllm-project/vllm-omni. Write MiniMax H3 video generation prompts for T2VA, I2VA, FL2VA, L2VA, and Ref2VA. Use when rewriting multimodal requests into H3 prompt structures, composing integratedmultimodaldescription, overallsoundscape, and nondiegeticmusic, aligning keyframes, or defining reference labels for images, videos, and audio.

Its SKILL.md is about 740 tokens, which your agent loads only when the skill is triggered. The skill folder holds 6 other files, including reference files (for example `agents/openai.yaml`, `references/base-format.md` and `references/portable-workflow.md`).

It sits in Media & Creative, covering AI video generation. It works with MiniMax. The repository describes itself as: A framework for efficient model inference with omni-modality models. The licence is Apache-2.0.

When your agent uses it

  • Rewriting multimodal requests into H3 prompt structures
  • Composing integratedmultimodaldescription
  • Overallsoundscape
  • Nondiegeticmusic

Example prompts

  • “/h3-prompt-writing”

Workflow steps

5 steps, taken from the first numbered list in SKILL.md.

  1. Read portable-workflow.md for scope, available tools, durable artifacts, and generation handoff. For prompt-only work, no generation tools…
  2. Identify the input mode: T2VA, I2VA, FL2VA, L2VA, or full-reference Ref2VA.
  3. For base text/keyframe modes, read references/base-format.md and follow its final prompt structure.
  4. For full-reference mode, read references/ref-format.md and follow its six-section rewrite format.
  5. Preserve the exact field names, section order, labels, and timing notation from the selected guide.

What it can do on your machine

Read from SKILL.md and the folder at commit 096988d. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    No scripts in the folder and no shell commands in SKILL.md.

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

H3 Prompt Writing loads about 744 tokens when it runs, and up to ~4.8k if it reads all its reference files. Until then it costs about 84 tokens; SKILL.md has 300 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~84
When it runs · the whole SKILL.md, loaded when a task matches
~744
With references · SKILL.md plus every file in references/, read only if the agent opens them
~4.8k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from vllm-project/vllm-omni at commit 096988d, republished under its Apache-2.0 licence (© vllm-project). 300 words, ~744 tokens.

Download SKILL.mdSave it as .claude/skills/h3-prompt-writing/SKILL.md (or your agent's skills folder). This skill also uses 4 other files; get the full folder from GitHub.
name
h3-prompt-writing
description
Write MiniMax H3 video generation prompts for T2VA, I2VA, FL2VA, L2VA, and Ref2VA. Use when rewriting multimodal requests into H3 prompt structures, composing integrated_multimodal_description, overall_soundscape, and non_diegetic_music, aligning keyframes, or defining reference labels for images, videos, and audio.
metadata.compatibility
Portable to any agent that can read local files — no external API calls, MiniMax Hub tools, or proprietary runtime required. The agents/openai.yaml file only…

H3 Prompt Writing

Workflow

  1. Read portable-workflow.md for scope, available tools, durable artifacts, and generation handoff. For prompt-only work, no generation tools are required.
  2. Identify the input mode: T2VA, I2VA, FL2VA, L2VA, or full-reference Ref2VA.
  3. For base text/keyframe modes, read references/base-format.md and follow its final prompt structure.
  4. For full-reference mode, read references/ref-format.md and follow its six-section rewrite format.
  5. Preserve the exact field names, section order, labels, and timing notation from the selected guide.

Base Modes

  • T2VA: build the full audiovisual timeline from text.
  • I2VA: start from the first frame and develop forward from it.
  • FL2VA: describe the continuous path between the first and last frames.
  • L2VA: infer a plausible opening and converge to the supplied last frame.

Use integrated_multimodal_description, overall_soundscape, and non_diegetic_music in the order shown in references/base-format.md.

Full-Reference Mode

Ref2VA rewrites use subject_definitions, summary, retention_analysis, detailed_description, overall_soundscape, and non_diegetic_music in that order. Reference labels stay consistent across all sections.

Read references/ref-format.md for label rules and retention analysis.

Output Rules

  • Write rewrite sections in English; preserve dialogue, lyrics, and visible scene text in their original language.
  • Describe each shot by composition, subjects, environment, actions, camera, sound, and the exact point where referenced content appears.
  • Avoid plot summaries, unresolved reference labels, and timing that does not match the requested duration.

Tips for Better Results

  • Match the description to the effective duration of the current clip or generation window. The upstream guides describe 4–15-second clips; longer work requires a verified backend extension or an explicit assembly plan, as described in the portable workflow.
  • Keep reference labels consistent (e.g. <Picture 1>, <Video 1>, <Audio 1>) across every section.
  • Prefer concrete visual and audio details over abstract words like "cinematic" or "beautiful".
  • When using keyframes (I2VA / FL2VA / L2VA), clearly state how the first and/or last frame connects to the timeline.

© vllm-project, Apache-2.0. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

SKILL.md and 4 other files (references) in .agents/skills/h3-prompt-writing of vllm-project/vllm-omni.

  • SKILL.md
  • agents/openai.yaml
  • references/base-format.md
  • references/portable-workflow.md
  • references/ref-format.md

Open the folder on GitHubat commit 096988d

Used in 6 other repositories

We found 8 copies of this SKILL.md (exact, near-identical or edited) in other folders, from 6 other GitHub owners. This page covers the copy in vllm-project/vllm-omni, which our catalogue first saw on October 8, 2026.

Compare with similar skills

H3 Prompt Writing next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

H3 Prompt Writing compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
H3 Prompt Writing this skillvllm-project/vllm-omni7.1k6 repos~744Automated safety check: PassApache-2.0
Minimax H3 Video Reversegnipbao/minimax-h3-video-reverse-skill230—~2.2kAutomated safety check: PassMIT
Minimax H3 Colabkillkli/minimax-h3-colab-skill110—~743Automated safety check: PassNone
Handdrawn Live Video Generatortl2012tl/comfyUI-llama-TE2414 repos~1.5kAutomated safety check: PassNone
ComfyUI Local DriverSlavaSexton/ComfyUI-Agent-Kit105—~12kAutomated safety check: PassApache-2.0
Minimax H3 Text Video Promptunknowlei/minimax-h3-opencode-skills122—~2.3kAutomated safety check: PassMIT

Similar skills

  • Minimax H3 Video Reverse

    gnipbao/minimax-h3-video-reverse-skill

    先复述参考视频内容与叙事逻辑,和用户确认理解后,再反推 MiniMax H3 文生视频或逐镜图生视频提示词。Use when the user asks to reverse-engineer a video into prompts, reconstruct shots, extract first/end-frame prompts, adapt a reference to H3, or…

    230 GitHub stars~2.2k tokensUpdated 8 days ago
    Media & CreativeAuto-check passed
  • Minimax H3 Colab

    killkli/minimax-h3-colab-skill

    Create short MiniMax H3 reference-to-video clips from one or more local images through Google Colab CLI.

    110 GitHub stars~743 tokensUpdated 14 days ago
    Media & CreativeAuto-check passed
  • Handdrawn Live Video Generator

    tl2012tl/comfyUI-llama-TE

    For creators making surreal short videos that blend rough glowing hand-drawn animation with live-action spaces.

    241 GitHub starsUsed in 4 repos~1.5k tokens
    Media & CreativeAuto-check passed
  • ComfyUI Local Driver

    SlavaSexton/ComfyUI-Agent-Kit

    Drives a local ComfyUI install over its HTTP API to generate and edit images, video and audio, with per-model prompt recipes and workflow guidance.

    105 GitHub stars~12k tokensUpdated 1 mo ago
    Media & CreativeAuto-check passed
  • Minimax H3 Text Video Prompt

    unknowlei/minimax-h3-opencode-skills

    Downstream MiniMax H3 specialist for professional text-to-video (T2VA) prompts using the official three-field format.

    122 GitHub stars~2.3k tokensUpdated 2 mo ago
    Media & CreativeAuto-check passed
  • Minimax H3 Creative Director

    unknowlei/minimax-h3-opencode-skills

    Primary mandatory entrypoint for every MiniMax H3 video-generation request.

    122 GitHub stars~2.8k tokensUpdated 2 mo ago
    Media & CreativeAuto-check passed

More from vllm-project/vllm-omni

All 20 skills in this repo
  • Diffusion Perf Opt

    vllm-project/vllm-omni

    Diagnose and optimize vLLM Omni diffusion workloads, especially Wan/Qwen/Flux-style image and video generation.

    7.1k GitHub stars~7.5k tokensUpdated today
    Auto-check passed
  • Precheck PR

    vllm-project/vllm-omni

    Self-check your branch before creating a PR — catch dead code, prevent new model-specific Python examples, verify accuracy/perf claims, validate PR title format, and confirm merge readiness.

    7.1k GitHub stars~1.4k tokensUpdated today
    Auto-check passed
  • Quantization

    vllm-project/vllm-omni

    Work on vLLM-Omni quantization for diffusion, autoregressive, omni, or multi-stage models.

    7.1k GitHub stars~1.4k tokensUpdated today
    Auto-check passed
  • Review PR

    vllm-project/vllm-omni

    Review pull requests and local branches for vllm-project/vllm-omni with a frozen snapshot, module-design ownership, feature-design overlays, targeted validation, and concise evidence-backed findings.

    7.1k GitHub stars~3.8k tokensUpdated today
    Auto-check passed
  • Add Diffusion Model

    vllm-project/vllm-omni

    Add a new diffusion model (text-to-image, text-to-video, image-to-video, text-to-audio, image editing) to vLLM-Omni, including native non-Diffusers ports, reference-parity validation, Cache-DiT…

    7.1k GitHub stars~7k tokensUpdated today
    Auto-check passed
  • Add Recipe

    vllm-project/vllm-omni

    Add or update an in-repository vLLM-Omni model recipe with verified task, input, output, hardware, command, feature, and validation contracts.

    7.1k GitHub stars~1.4k tokensUpdated today
    Auto-check passed

Works with

Questions about H3 Prompt Writing

What does H3 Prompt Writing do?

Write MiniMax H3 video generation prompts for T2VA, I2VA, FL2VA, L2VA, and Ref2VA. H3 Prompt Writing is an agent skill from vllm-project/vllm-omni. Write MiniMax H3 video generation prompts for T2VA, I2VA, FL2VA, L2VA, and Ref2VA.

When should I use H3 Prompt Writing?

H3 Prompt Writing fits situations like: rewriting multimodal requests into H3 prompt structures; composing integratedmultimodaldescription; overallsoundscape; nondiegeticmusic.

How do I install H3 Prompt Writing in Claude Code?

Run `npx skills add vllm-project/vllm-omni --skill h3-prompt-writing -a claude-code`. Or copy the skill folder (.agents/skills/h3-prompt-writing in vllm-project/vllm-omni) into .claude/skills/h3-prompt-writing in your project. Claude Code loads it when a task matches its description.

How do I install H3 Prompt Writing in Codex?

Run `npx skills add vllm-project/vllm-omni --skill h3-prompt-writing -a codex`. Or copy the skill folder (.agents/skills/h3-prompt-writing in vllm-project/vllm-omni) into .agents/skills/h3-prompt-writing in your project. Codex loads it when a task matches its description.

Can I use H3 Prompt Writing in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add vllm-project/vllm-omni --skill h3-prompt-writing -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/h3-prompt-writing, .gemini/skills/h3-prompt-writing, .github/skills/h3-prompt-writing and .opencode/skills/h3-prompt-writing in your project.

What does H3 Prompt Writing need to run?

SKILL.md names no scripts, command-line tools or credentials: H3 Prompt Writing is instructions for the agent only.

Does H3 Prompt Writing access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is H3 Prompt Writing safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does H3 Prompt Writing use?

H3 Prompt Writing is published under the Apache-2.0 licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does H3 Prompt Writing use?

About 744 tokens (SKILL.md is roughly 3k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full. Its references folder adds about 4.1k tokens, read only when the agent opens those files.

What are the alternatives to H3 Prompt Writing?

Skills that share tags, products or a category with H3 Prompt Writing: Minimax H3 Video Reverse (gnipbao/minimax-h3-video-reverse-skill, 230 stars), Minimax H3 Colab (killkli/minimax-h3-colab-skill, 110 stars), Handdrawn Live Video Generator (tl2012tl/comfyUI-llama-TE, 241 stars) and ComfyUI Local Driver (SlavaSexton/ComfyUI-Agent-Kit, 105 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains H3 Prompt Writing?

vllm-project (a GitHub organization) maintains it in vllm-project/vllm-omni, which has 7,119 GitHub stars. The repository holds 20 skills in this directory. The repository was last updated on October 11, 2026.

Source: vllm-project/vllm-omni on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.