Agent skill

Minimax H3 Text Video Prompt

by unknowlei in unknowlei/minimax-h3-opencode-skills

Downstream MiniMax H3 specialist for professional text-to-video (T2VA) prompts using the official three-field format.

MITAuto-check passedMedia & Creative

Install Minimax H3 Text Video Prompt

skills CLI
$ npx skills add unknowlei/minimax-h3-opencode-skills --skill minimax-h3-text-video-prompt -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install unknowlei/minimax-h3-opencode-skills minimax-h3-text-video-prompt --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/unknowlei/minimax-h3-opencode-skills.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/minimax-h3-text-video-prompt .claude/skills/minimax-h3-text-video-prompt && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
minimax-h3-text-video-prompt
GitHub stars
122
Token cost
~2.3k tokens
SKILL.md length
1,172 words
Files
3 (incl. references)
Skills in repo
6
Repo updated
First seen
Licence
MIT

At a glance

Downstream MiniMax H3 specialist for professional text-to-video (T2VA) prompts using the official three-field format.

  • Works in 9 steps: Extract or reasonably infer duration,… → Run the interactive direction check when… → Budget actions and cuts to fit the… → …
  • Tasks that involve AI video generation
  • SKILL.md covers Routing Contract, Official Format Authority, Confirmed Multishot Handoff and Mandatory Format, plus 6 more sections
  • Instructions only: no scripts, shell commands, URLs or credentials in SKILL.md

What it does

Minimax H3 Text Video Prompt is an agent skill from unknowlei/minimax-h3-opencode-skills. Downstream MiniMax H3 specialist for professional text-to-video (T2VA) prompts using the official three-field format. Use after minimax-h3-creative-director routes a request with no image, video, or audio reference asset, or when this skill is explicitly invoked for a text-only idea, script, or storyboard requiring an audiovisual timeline and Chinese translation.

Its SKILL.md is about 2.3k tokens, which your agent loads only when the skill is triggered. The skill folder holds 4 other files, including reference files (for example `agents/openai.yaml` and `references/official-rules.md`).

It sits in Media & Creative, covering AI video generation, Comics and storyboards and Translation. It works with MiniMax. The repository describes itself as: OpenCode skill suite for MiniMax H3 directing, routing, multishot planning, prompt generation, and review. The licence is MIT.

When your agent uses it

  • Tasks that involve AI video generation
  • Tasks that involve Comics and storyboards
  • Tasks that involve Translation

Example prompts

  • “/minimax-h3-text-video-prompt”

Workflow steps

9 steps, taken from the first numbered list in SKILL.md.

  1. Extract or reasonably infer duration, aspect ratio, visual style, subjects, setting, action arc, dialogue, sound, and desired ending. Keep…
  2. Run the interactive direction check when required. Never invent dialogue, lyrics, or visible text.
  3. Budget actions and cuts to fit the requested duration. Prefer one coherent action arc over many incomplete events.
  4. Establish style, composition, subject appearance and position, environment, lighting, and initial state at [Shot 1].
  5. Describe visible state changes in playback order: initial state -> action onset -> continuous development -> result or reaction.
  6. Add cuts only when they introduce a meaningful change in subject, space, state, viewpoint, or time. Otherwise use camera movement.
  7. Assign stable (S1), (S2) IDs only to actual vocal sources. Preserve user-provided words and punctuation inside [Language] ....
  8. Separate synchronized audiovisual events, overall physical soundscape, and audience-only music.
  9. Perform the quality checks below, then return the required bilingual output.

What it can do on your machine

Read from SKILL.md and the folder at commit 2e2096f. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    No scripts in the folder and no shell commands in SKILL.md.

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Minimax H3 Text Video Prompt loads about 2.3k tokens when it runs, and up to ~3.1k if it reads all its reference files. Until then it costs about 99 tokens; SKILL.md has 1,172 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~99
When it runs · the whole SKILL.md, loaded when a task matches
~2.3k
With references · SKILL.md plus every file in references/, read only if the agent opens them
~3.1k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from unknowlei/minimax-h3-opencode-skills at commit 2e2096f, republished under its MIT licence (© unknowlei). 1,172 words, ~2,274 tokens.

Download SKILL.mdSave it as .claude/skills/minimax-h3-text-video-prompt/SKILL.md (or your agent's skills folder). This skill also uses 2 other files; get the full folder from GitHub.
name
minimax-h3-text-video-prompt
description
Downstream MiniMax H3 specialist for professional text-to-video (T2VA) prompts using the official three-field format. Use after minimax-h3-creative-director routes a request with no image, video, or audio reference asset, or when this skill is explicitly invoked for a text-only idea, script, or storyboard requiring an audiovisual timeline and Chinese translation.

MiniMax H3 Text-to-Video Prompt

Turn a text-only video concept into a precise MiniMax H3 T2VA timeline. Write observable audiovisual instructions rather than a plot summary or a keyword pile.

Routing Contract

New MiniMax H3 requests should enter through minimax-h3-creative-director. If invoked directly, apply the official authority below and proceed without requiring a second routing round. Redirect to the keyframe or full-reference specialist when supplied assets materially control the result.

Official Format Authority

Before drafting, read ../h3-prompt-writing/SKILL.md and ../h3-prompt-writing/references/base-en.txt. Treat those official MiniMax files as the canonical prompt-format specification. If this skill or its local references conflict with them, follow the official files.

Read references/official-rules.md before drafting.

Confirmed Multishot Handoff

When the director supplies a confirmed multishot_plan, treat its shot count, timing, content, framing, performance, camera, transitions, sound, and continuity decisions as already answered. Do not ask those questions again. Map every confirmed shot into integrated_multimodal_description without dropping or silently changing choices. Reopen a decision only when the plan is physically impossible, internally contradictory, or exceeds the effective duration.

Mandatory Format

Always return the official three-field prompt. Never return a free-form natural-language paragraph, keyword list, abbreviated prompt, or alternate schema as the final English prompt, even when the user's input is informal or brief. Expand the request into the required structure.

Interactive Direction Check

Before drafting, use the host's structured-choice tool when any of these conditions applies. This is mandatory, not optional:

  • The brief is short, generic, or contains only a subject and a basic action.
  • Two or more of duration, scene/event progression, visual style, camera/edit rhythm, dialogue/voice, sound/music, or ending are missing.
  • A vague phrase has multiple materially different interpretations.
  • The user says to improvise, decide freely, surprise them, or let the AI choose.
  • Duration, visual treatment, action intensity, camera rhythm, or sound direction would substantially change the result.

In OpenCode, call the built-in question tool rather than printing Markdown options. In Codex, use the available structured user-input tool; if none is available, ask one concise plain-text question.

Ask as many related questions as the current decision stage requires; normally 1-5, but three is not a per-call, per-turn, or per-session ceiling. A strict yes/no question may contain exactly two options; every other choice question must contain at least five materially different, feasible options. Put the context-specific recommendation first and suffix its label with (Recommended). Give every option a one-sentence consequence. Do not add Other; OpenCode supplies custom input automatically. Use multiple selection only when choices can genuinely be combined. Before every question tool call, count the options in the actual payload. If any non-binary choice has fewer than five, do not submit it; expand it with meaningful alternatives or make it open-ended. Do not create near-duplicates merely to reach five.

Prefer questions about:

  • Visual medium and style: live-action cinematic, animation, stylized commercial, documentary, and similar directions
  • Narrative/action intensity: restrained, balanced, dynamic
  • Camera rhythm: one continuous shot, limited cinematic cuts, fast montage
  • Sound direction: natural ambience only, restrained score, prominent audiovisual design

For a sparse or delegated brief, ask enough high-impact questions in one batch to establish the current stage, normally 3-5, rather than silently completing the concept. Do not ask about minor details after those answers establish the direction. Skip questions only when the user explicitly says not to ask; a request to "freely create" or "improvise" requires questions rather than skipping them.

Workflow

  1. Extract or reasonably infer duration, aspect ratio, visual style, subjects, setting, action arc, dialogue, sound, and desired ending. Keep the official online target within 4-15 seconds, 24 FPS, and 7000 prompt characters; obey narrower local-workflow limits when known.
  2. Run the interactive direction check when required. Never invent dialogue, lyrics, or visible text.
  3. Budget actions and cuts to fit the requested duration. Prefer one coherent action arc over many incomplete events.
  4. Establish style, composition, subject appearance and position, environment, lighting, and initial state at [Shot 1].
  5. Describe visible state changes in playback order: initial state -> action onset -> continuous development -> result or reaction.
  6. Add cuts only when they introduce a meaningful change in subject, space, state, viewpoint, or time. Otherwise use camera movement.
  7. Assign stable (S1), (S2) IDs only to actual vocal sources. Preserve user-provided words and punctuation inside <d>[Language] ...</d>.
  8. Separate synchronized audiovisual events, overall physical soundscape, and audience-only music.
  9. Perform the quality checks below, then return the required bilingual output.
Show full SKILL.md (449 more words)Show less

Required Prompt Structure

Produce exactly these three fields in this order:

text
integrated_multimodal_description: [Shot 1] ...

overall_soundscape: ...

non_diegetic_music: ...

Do not add an image-alignment instruction for T2VA.

Output Contract

Return:

  1. ### English Prompt
  2. One text code block containing only the final English prompt
  3. ### 中文翻译
  4. A complete Chinese translation outside any code block

Language Policy

Write every instruction in the English prompt in English except content that must remain exact:

  • Dialogue, voiceover, speech, singing, and lyrics inside <d>[Language] ...</d>
  • Text visibly present in the video, including signs, captions, subtitles, labels, interfaces, titles, logos, and messages
  • Proper nouns whose exact spelling matters, including personal names, place names, brands, organizations, product/work titles, account names, and handles
  • Literal identifiers such as filenames, URLs, model names, reference labels, and control tokens

Do not use Chinese or another non-English language for scene description, style, composition, appearance, lighting, action, camera, editing, sound design, or music unless the user explicitly requires that literal wording to appear or be spoken in the video.

Translate every descriptive part of the completed prompt into Chinese after the English code block, including the meaning of dialogue, lyrics, and visible text. Keep structural identifiers such as field names, [Shot N], timestamps, (Sx), and control tags visible alongside their Chinese meanings so the translation can be mapped line by line. Preserve exact proper nouns and source literals, adding a Chinese explanation where useful. The Chinese section is explanatory and is not an executable MiniMax prompt.

If assumptions materially help the user, place a short ### 采用的假设 section before the English prompt. Do not put explanations or alternatives inside the English code block.

Keep the executable English prompt within the official 7000-character maximum. Use the detail required for professional control while removing repetition, contradictions, and details that cannot affect visible or audible output.

Quality Gate

Before answering, verify that:

  • [Shot 1] has no timestamp; later cut times are strictly increasing and inside the duration.
  • Every shot establishes the current composition before describing changes.
  • Actions have a physically understandable start, progression, and result.
  • Camera terms are mechanically correct and written as natural English actions.
  • Dialogue is exact, speaker IDs are stable, and voiceover cannot cause unintended lip movement.
  • Dialogue length is feasible for its shot; cross-cut speech explicitly continues as a J-cut or L-cut when applicable.
  • Every required title, sign, logo, caption, subtitle, slogan, or interface label is quoted exactly.
  • overall_soundscape does not repeat dialogue, singing, or diegetic music.
  • non_diegetic_music specifies instrumentation, tempo/rhythm, and dynamics, or is N/A.
  • The prompt does not contradict itself about one-take versus cuts or music versus N/A.
  • The final English prompt is no more than 7000 characters.
  • The Chinese translation covers every English instruction and every literal verbal or visible element without introducing new creative content.

© unknowlei, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

SKILL.md and 2 other files (references) in skills/minimax-h3-text-video-prompt of unknowlei/minimax-h3-opencode-skills.

  • SKILL.md
  • agents/openai.yaml
  • references/official-rules.md

Open the folder on GitHubat commit 2e2096f

Compare with similar skills

Minimax H3 Text Video Prompt next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Minimax H3 Text Video Prompt compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Minimax H3 Text Video Prompt this skillunknowlei/minimax-h3-opencode-skills122—~2.3kAutomated safety check: PassMIT
Storyboardgodot-fun/gai183—~1.7kAutomated safety check: PassMIT
MiniMax H3 Video DirectorTFboy1/oh-my-minimaxh3-director143—~2.4kAutomated safety check: PassMIT
Hong Kong Comic Fighter for H3karuvanan/MiniMax-H3-Director-Cut-Studio132—~4.4kAutomated safety check: PassCustom licence
Minimax H3 Video Reversegnipbao/minimax-h3-video-reverse-skill230—~2.2kAutomated safety check: PassMIT
H3 Cinematic Directorjtydhr88/ComfyTV1.1k—~2.7kAutomated safety check: PassMIT

Similar skills

  • Storyboard

    godot-fun/gai

    Turns user materials into a shot-by-shot storyboard with bilingual narration (Chinese + English VO, both required) and AI video prompts per shot.

    183 GitHub stars~1.7k tokensUpdated yesterday
    Media & CreativeAuto-check passed
  • MiniMax H3 Video Director

    TFboy1/oh-my-minimaxh3-director

    Turns a script into a storyboard, assigns MiniMax H3 workflows in ComfyUI, monitors batch generation and builds a Jianying draft of the finished video.

    143 GitHub stars~2.4k tokensUpdated 1 mo ago
    Media & CreativeAuto-check passed
  • Hong Kong Comic Fighter for H3

    karuvanan/MiniMax-H3-Director-Cut-Studio

    Special skill that turns loaded Hong Kong comic panels into photoreal MiniMax H3 martial-arts sequences with readable attack and defence action.

    132 GitHub stars~4.4k tokensUpdated today
    Media & CreativeAuto-check passed
  • Minimax H3 Video Reverse

    gnipbao/minimax-h3-video-reverse-skill

    先复述参考视频内容与叙事逻辑,和用户确认理解后,再反推 MiniMax H3 文生视频或逐镜图生视频提示词。Use when the user asks to reverse-engineer a video into prompts, reconstruct shots, extract first/end-frame prompts, adapt a reference to H3, or…

    230 GitHub stars~2.2k tokensUpdated 8 days ago
    Media & CreativeAuto-check passed
  • H3 Cinematic Director

    jtydhr88/ComfyTV

    Convert approved scripts, shot briefs, storyboards, keyframes, character sheets, scene assets, prop sheets, and sound briefs into director-level storyboards and production-ready MiniMax H3 prompts.

    1.1k GitHub stars~2.7k tokensUpdated 4 days ago
    Media & CreativeAuto-check passed
  • Scroll Promo Site Builder

    kangarooking/kangarooking-skills

    Create a scroll-controlled cinematic product website with rich motion (动效网站) from product materials, reference pages or videos, and brand assets.

    662 GitHub stars~2.6k tokensUpdated 1 mo ago
    Media & CreativeAuto-check passed

More from unknowlei/minimax-h3-opencode-skills

  • Minimax H3 Multishot Planner

    unknowlei/minimax-h3-opencode-skills

    Non-skippable planning-only MiniMax H3 subskill used by minimax-h3-creative-director before final prompt formatting.

    122 GitHub stars~2.2k tokensUpdated 2 mo ago
    Auto-check passed
  • Minimax H3 Keyframe Video Prompt

    unknowlei/minimax-h3-opencode-skills

    Narrow downstream MiniMax H3 specialist for pure I2VA, FL2VA, and L2VA boundary-frame prompts.

    122 GitHub stars~2.4k tokensUpdated 2 mo ago
    Auto-check passed
  • Minimax H3 Reference Video Prompt

    unknowlei/minimax-h3-opencode-skills

    Default downstream MiniMax H3 specialist for every image-based request unless the user explicitly declares boundary-only first/last frames with no reusable reference role.

    122 GitHub stars~3k tokensUpdated 2 mo ago
    Auto-check passed
  • Minimax H3 Creative Director

    unknowlei/minimax-h3-opencode-skills

    Primary mandatory entrypoint for every MiniMax H3 video-generation request.

    122 GitHub stars~2.8k tokensUpdated 2 mo ago
    Auto-check passed
  • Minimax H3 Prompt Reviewer

    unknowlei/minimax-h3-opencode-skills

    Downstream MiniMax H3 specialist that audits, repairs, and rewrites T2VA, I2VA, FL2VA, L2VA, and full-reference prompts into an official structured format.

    122 GitHub stars~2.1k tokensUpdated 2 mo ago
    Auto-check passed

Works with

Questions about Minimax H3 Text Video Prompt

What does Minimax H3 Text Video Prompt do?

Downstream MiniMax H3 specialist for professional text-to-video (T2VA) prompts using the official three-field format. Minimax H3 Text Video Prompt is an agent skill from unknowlei/minimax-h3-opencode-skills. Downstream MiniMax H3 specialist for professional text-to-video (T2VA) prompts using the official three-field format.

When should I use Minimax H3 Text Video Prompt?

Minimax H3 Text Video Prompt fits situations like: tasks that involve AI video generation; tasks that involve Comics and storyboards; tasks that involve Translation.

How do I install Minimax H3 Text Video Prompt in Claude Code?

Run `npx skills add unknowlei/minimax-h3-opencode-skills --skill minimax-h3-text-video-prompt -a claude-code`. Or copy the skill folder (skills/minimax-h3-text-video-prompt in unknowlei/minimax-h3-opencode-skills) into .claude/skills/minimax-h3-text-video-prompt in your project. Claude Code loads it when a task matches its description.

How do I install Minimax H3 Text Video Prompt in Codex?

Run `npx skills add unknowlei/minimax-h3-opencode-skills --skill minimax-h3-text-video-prompt -a codex`. Or copy the skill folder (skills/minimax-h3-text-video-prompt in unknowlei/minimax-h3-opencode-skills) into .agents/skills/minimax-h3-text-video-prompt in your project. Codex loads it when a task matches its description.

Can I use Minimax H3 Text Video Prompt in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add unknowlei/minimax-h3-opencode-skills --skill minimax-h3-text-video-prompt -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/minimax-h3-text-video-prompt, .gemini/skills/minimax-h3-text-video-prompt, .github/skills/minimax-h3-text-video-prompt and .opencode/skills/minimax-h3-text-video-prompt in your project.

What does Minimax H3 Text Video Prompt need to run?

SKILL.md names no scripts, command-line tools or credentials: Minimax H3 Text Video Prompt is instructions for the agent only.

Does Minimax H3 Text Video Prompt access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is Minimax H3 Text Video Prompt safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Minimax H3 Text Video Prompt use?

Minimax H3 Text Video Prompt is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Minimax H3 Text Video Prompt use?

About 2.3k tokens (SKILL.md is roughly 9.1k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full. Its references folder adds about 826 tokens, read only when the agent opens those files.

What are the alternatives to Minimax H3 Text Video Prompt?

Skills that share tags, products or a category with Minimax H3 Text Video Prompt: Storyboard (godot-fun/gai, 183 stars), MiniMax H3 Video Director (TFboy1/oh-my-minimaxh3-director, 143 stars), Hong Kong Comic Fighter for H3 (karuvanan/MiniMax-H3-Director-Cut-Studio, 132 stars) and Minimax H3 Video Reverse (gnipbao/minimax-h3-video-reverse-skill, 230 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Minimax H3 Text Video Prompt?

unknowlei (a GitHub user) maintains it in unknowlei/minimax-h3-opencode-skills, which has 122 GitHub stars. The repository holds 6 skills in this directory. The repository was last updated on August 9, 2026.

Source: unknowlei/minimax-h3-opencode-skills on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.