Agent skill

Minimax H3 Keyframe Video Prompt

by unknowlei in unknowlei/minimax-h3-opencode-skills

Narrow downstream MiniMax H3 specialist for pure I2VA, FL2VA, and L2VA boundary-frame prompts.

MITAuto-check passedMedia & Creative

Install Minimax H3 Keyframe Video Prompt

skills CLI
$ npx skills add unknowlei/minimax-h3-opencode-skills --skill minimax-h3-keyframe-video-prompt -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install unknowlei/minimax-h3-opencode-skills minimax-h3-keyframe-video-prompt --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/unknowlei/minimax-h3-opencode-skills.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/minimax-h3-keyframe-video-prompt .claude/skills/minimax-h3-keyframe-video-prompt && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
minimax-h3-keyframe-video-prompt
GitHub stars
122
Token cost
~2.4k tokens
SKILL.md length
1,208 words
Files
3 (incl. references)
Skills in repo
6
Repo updated
First seen
Licence
MIT

At a glance

Narrow downstream MiniMax H3 specialist for pure I2VA, FL2VA, and L2VA boundary-frame prompts.

  • Works in 8 steps: Confirm the boundary role, exact integer… → Inspect supplied images when available.… → Define the motion path before writing… → …
  • Media & Creative work in your project
  • SKILL.md covers Routing Contract, Official Format Authority, Confirmed Multishot Handoff and Mandatory Format, plus 8 more sections
  • Instructions only: no scripts, shell commands, URLs or credentials in SKILL.md

What it does

Minimax H3 Keyframe Video Prompt is an agent skill from unknowlei/minimax-h3-opencode-skills. Narrow downstream MiniMax H3 specialist for pure I2VA, FL2VA, and L2VA boundary-frame prompts. Use only after minimax-h3-creative-director verifies that the user explicitly declared the images as literal first/last frames and that they have no character, identity, person, object, costume, scene, style, voice, action, camera, or other reusable reference responsibility. Otherwise use full-reference consistency, even when first/last-frame semantics are desired.

Its SKILL.md is about 2.4k tokens, which your agent loads only when the skill is triggered. The skill folder holds 4 other files, including reference files (for example `agents/openai.yaml` and `references/keyframe-rules.md`).

It sits in Media & Creative. It works with MiniMax. The repository describes itself as: OpenCode skill suite for MiniMax H3 directing, routing, multishot planning, prompt generation, and review. The licence is MIT.

When your agent uses it

  • Media & Creative work in your project

Example prompts

  • “/minimax-h3-keyframe-video-prompt”

Workflow steps

8 steps, taken from the first numbered list in SKILL.md.

  1. Confirm the boundary role, exact integer duration, and intended final action or transformation. Keep the official online target within…
  2. Inspect supplied images when available. Record identity, clothing, pose, object states, composition, camera angle, lighting, and spatial…
  3. Define the motion path before writing prose.
  4. Keep subject identity and persistent attributes consistent unless the user explicitly requests transformation.
  5. Prefer a single continuous transition for FL2VA and usually for L2VA. Two boundary images do not automatically create a cut. Add cuts only…
  6. Write the exact alignment instruction, then the three core fields.
  7. Validate that the target boundary frame is reached at the exact endpoint rather than early.
  8. Return the required bilingual output.

What it can do on your machine

Read from SKILL.md and the folder at commit 2e2096f. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    No scripts in the folder and no shell commands in SKILL.md.

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Minimax H3 Keyframe Video Prompt loads about 2.4k tokens when it runs, and up to ~3.2k if it reads all its reference files. Until then it costs about 124 tokens; SKILL.md has 1,208 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~124
When it runs · the whole SKILL.md, loaded when a task matches
~2.4k
With references · SKILL.md plus every file in references/, read only if the agent opens them
~3.2k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from unknowlei/minimax-h3-opencode-skills at commit 2e2096f, republished under its MIT licence (© unknowlei). 1,208 words, ~2,386 tokens.

Download SKILL.mdSave it as .claude/skills/minimax-h3-keyframe-video-prompt/SKILL.md (or your agent's skills folder). This skill also uses 2 other files; get the full folder from GitHub.
name
minimax-h3-keyframe-video-prompt
description
Narrow downstream MiniMax H3 specialist for pure I2VA, FL2VA, and L2VA boundary-frame prompts. Use only after minimax-h3-creative-director verifies that the user explicitly declared the images as literal first/last frames and that they have no character, identity, person, object, costume, scene, style, voice, action, camera, or other reusable reference responsibility. Otherwise use full-reference consistency, even when first/last-frame semantics are desired.

MiniMax H3 Keyframe Video Prompt

Design a continuous visual path from or toward concrete boundary frames. Treat a boundary image as an actual target frame, not merely a loose style reference.

Routing Contract

New MiniMax H3 requests should enter through minimax-h3-creative-director. Use this skill only for explicitly declared boundary-only images. If the user did not explicitly say first frame, last frame, starting frame, ending frame, or an equivalent literal boundary role, use minimax-h3-reference-video-prompt. If any image also defines or preserves a character, identity, person, object, costume, scene, style, voice, action, camera, composition, or other reusable trait, use full-reference even if the user also calls it a first or last frame.

Official Format Authority

Before drafting, read ../h3-prompt-writing/SKILL.md and ../h3-prompt-writing/references/base-en.txt. Treat those official MiniMax files as the canonical prompt-format specification. If this skill or its local references conflict with them, follow the official files.

Read references/keyframe-rules.md before drafting.

Confirmed Multishot Handoff

When the director supplies a confirmed multishot_plan, treat its shot count, timing, content, framing, performance, camera, transitions, sound, and continuity decisions as already answered. Do not ask those questions again. Map every confirmed shot into the official keyframe prompt while preserving exact boundary-frame alignment. Reopen a decision only when the plan is physically impossible, contradicts a boundary image, or exceeds the effective duration.

Mandatory Format

Always return the exact mode-specific picture-alignment instruction followed by the official three fields. Never return a free-form natural-language paragraph, keyword list, shortened prompt, or alternate schema as the final English prompt.

Interactive Direction Check

Before drafting, use the host's structured-choice tool whenever the eligible pure-keyframe brief is sparse, the intended transformation is underspecified, two or more of duration, motion path, visual treatment, camera strategy, dialogue/sound, or ending behavior are missing, or the user delegates creative direction to the AI.

In OpenCode, call the built-in question tool rather than printing Markdown options. In Codex, use the available structured user-input tool; if unavailable, ask one concise plain-text question.

Ask as many related questions as the current decision stage requires; normally 1-5, but three is not a per-call, per-turn, or per-session ceiling. A strict yes/no question may contain exactly two options; every other choice question must contain at least five materially different, feasible options. Put the context-specific recommendation first with (Recommended) in its label, explain each consequence in one sentence, omit Other, and enable multiple selection only for genuinely combinable choices. Before every question tool call, count the options in the actual payload. If any non-binary choice has fewer than five, do not submit it; expand it with meaningful alternatives or make it open-ended. Do not create near-duplicates merely to reach five.

Prioritize:

  • Boundary role when one image is ambiguous: first frame or last frame
  • Motion strength: subtle natural motion, clear narrative action, dramatic transformation
  • Camera strategy: continuous single shot, restrained cuts, dynamic montage
  • Landing behavior: exact stable hold, arrive on the final instant, transition through the frame

For a sparse or delegated brief, ask enough high-impact questions in one batch to establish the current stage, normally 3-5. Do not treat "improvise" as permission to skip questions. Skip only when the user explicitly prohibits questions, then disclose the chosen direction.

Select the Mode

  • Explicit first-frame-only image with no reusable reference role -> I2VA.
  • Explicit first-and-last-frame pair with no reusable reference role -> FL2VA.
  • Explicit last-frame-only image with no reusable reference role -> L2VA.

If boundary wording is absent, ambiguous, or merely inferred, do not ask the user to confirm this specialist. Route directly to minimax-h3-reference-video-prompt.

If the same boundary image or any other asset must also control identity, character, object, action, style, scene, sound, editing, or pacing, use minimax-h3-reference-video-prompt and express the first/last-frame relationship inside the full-reference prompt. Do not mix API first_frame/last_frame roles with reference_* roles.

Workflow

  1. Confirm the boundary role, exact integer duration, and intended final action or transformation. Keep the official online target within 4-15 seconds, 24 FPS, and 7000 prompt characters; obey narrower local-workflow limits when known.
  2. Inspect supplied images when available. Record identity, clothing, pose, object states, composition, camera angle, lighting, and spatial relationships.
  3. Define the motion path before writing prose.
  4. Keep subject identity and persistent attributes consistent unless the user explicitly requests transformation.
  5. Prefer a single continuous transition for FL2VA and usually for L2VA. Two boundary images do not automatically create a cut. Add cuts only when explicitly requested or essential.
  6. Write the exact alignment instruction, then the three core fields.
  7. Validate that the target boundary frame is reached at the exact endpoint rather than early.
  8. Return the required bilingual output.
Show full SKILL.md (454 more words)Show less

Motion Paths

I2VA

Use: first-frame anchor -> action onset -> continuous development -> result or reaction.

FL2VA

Use: first-frame state -> observable intermediate changes -> progressively narrowing differences -> exact last-frame state.

Do not merely describe two static images. Explain how poses, objects, lighting, camera, and composition transform between them.

L2VA

Use: plausible preceding state -> explicit causal action -> gradual convergence -> exact final-frame landing.

Do not begin in the final state. Infer a compatible earlier state and show how it becomes the reference frame.

Required Prompt Structure

Place the mode-specific alignment instruction first, followed by one blank line and:

text
integrated_multimodal_description: [Shot 1] ...

overall_soundscape: ...

non_diegetic_music: ...

Use the exact official alignment forms in references/keyframe-rules.md.

Output Contract

Return ### English Prompt, one text code block containing only the complete English prompt, then ### 中文翻译 with a complete Chinese translation outside code blocks.

Language Policy

Write all descriptive and instructional prose in the English prompt in English. Preserve non-English source wording only when exactness is required for:

  • Dialogue, voiceover, speech, singing, or lyrics inside <d>[Language] ...</d>
  • Visible signs, captions, subtitles, labels, interfaces, titles, logos, or messages
  • Exact personal names, place names, brands, organizations, product/work titles, account names, and handles
  • Filenames, URLs, model names, picture labels, and control tokens

Do not use non-English wording for style, composition, appearance, lighting, actions, transitions, camera movement, sound design, or music unless it is literal content to be spoken or shown.

Translate every completed prompt instruction into Chinese after the English code block, including the meaning of dialogue, lyrics, and visible text. Keep field names, picture labels, [Shot N], timestamps, (Sx), and control tags visible alongside their Chinese explanations. Preserve exact proper nouns and source literals, adding Chinese meaning where useful. Treat the Chinese section as explanatory rather than executable.

Add ### 采用的假设 before the English prompt only when useful. Keep the executable English prompt within the official 7000-character maximum; choose detail according to transition complexity and remove repetition.

Quality Gate

  • The selected mode matches the supplied boundary frames.
  • The alignment instruction is exact and uses the effective duration with exactly two decimal places where required.
  • The initial state matches the first-frame image when one exists.
  • Intermediate changes are visible, causal, and feasible within the duration.
  • Identity, clothing, key objects, spatial relations, and persistent colors remain continuous.
  • The final pose, object state, camera angle, lighting, spacing, and composition land on the last frame when one exists.
  • The final frame is reached at the end, not held prematurely.
  • Dialogue fits its shot duration, and any cross-cut speech relationship is explicit.
  • Required visible text is quoted exactly.
  • The prompt does not contradict one-take versus cuts or music versus N/A.
  • The final English prompt is no more than 7000 characters.
  • Dialogue, sound, and camera follow the shared MiniMax H3 rules, and the Chinese section translates every prompt component without adding content.

© unknowlei, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

SKILL.md and 2 other files (references) in skills/minimax-h3-keyframe-video-prompt of unknowlei/minimax-h3-opencode-skills.

  • SKILL.md
  • agents/openai.yaml
  • references/keyframe-rules.md

Open the folder on GitHubat commit 2e2096f

Compare with similar skills

Minimax H3 Keyframe Video Prompt next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Minimax H3 Keyframe Video Prompt compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Minimax H3 Keyframe Video Prompt this skillunknowlei/minimax-h3-opencode-skills122—~2.4kAutomated safety check: PassMIT
AI Image Generation and Editingzhayujie/CowAgent47k—~1.3kAutomated safety check: PassMIT
Gif Sticker MakerHHU3637kr/skills1453 repos~1.3kAutomated safety check: PassMIT
Minimax H3 Video Reversegnipbao/minimax-h3-video-reverse-skill230—~2.2kAutomated safety check: PassMIT
H3 Cinematic Directorjtydhr88/ComfyTV1.1k—~2.7kAutomated safety check: PassMIT
Novel Storyboardeternityspring/shuohao-skills4.3k—~1.8kAutomated safety check: NotesApache-2.0

Similar skills

  • Generates or edits images from text prompts through a Python script that picks an image backend based on which API keys are configured.

    47k GitHub stars~1.3k tokensUpdated yesterday
    Media & CreativeAuto-check passed
  • Gif Sticker Maker

    HHU3637kr/skills

    Convert photos (people, pets, objects, logos) into 4 animated GIF stickers with captions.

    145 GitHub starsUsed in 3 repos~1.3k tokens
    Media & CreativeAuto-check passed
  • Minimax H3 Video Reverse

    gnipbao/minimax-h3-video-reverse-skill

    先复述参考视频内容与叙事逻辑,和用户确认理解后,再反推 MiniMax H3 文生视频或逐镜图生视频提示词。Use when the user asks to reverse-engineer a video into prompts, reconstruct shots, extract first/end-frame prompts, adapt a reference to H3, or…

    230 GitHub stars~2.2k tokensUpdated 9 days ago
    Media & CreativeAuto-check passed
  • H3 Cinematic Director

    jtydhr88/ComfyTV

    Convert approved scripts, shot briefs, storyboards, keyframes, character sheets, scene assets, prop sheets, and sound briefs into director-level storyboards and production-ready MiniMax H3 prompts.

    1.1k GitHub stars~2.7k tokensUpdated 4 days ago
    Media & CreativeAuto-check passed
  • Novel Storyboard

    eternityspring/shuohao-skills

    给 AI 短剧出分镜:三层结构——段(一次视频生成,≤15 秒)→ 分镜(段内 2–5 秒的剪切,认领剧本节拍) → 分镜图(每切一张关键帧:主分镜图钉 0.00 秒,子分镜图钉各自切点)。

    4.3k GitHub stars~1.8k tokensUpdated 3 days ago
    Media & CreativeAuto-check: notes
  • Minimax H3 Colab

    killkli/minimax-h3-colab-skill

    Create short MiniMax H3 reference-to-video clips from one or more local images through Google Colab CLI.

    110 GitHub stars~743 tokensUpdated 15 days ago
    Media & CreativeAuto-check passed

More from unknowlei/minimax-h3-opencode-skills

  • Minimax H3 Multishot Planner

    unknowlei/minimax-h3-opencode-skills

    Non-skippable planning-only MiniMax H3 subskill used by minimax-h3-creative-director before final prompt formatting.

    122 GitHub stars~2.2k tokensUpdated 2 mo ago
    Auto-check passed
  • Minimax H3 Reference Video Prompt

    unknowlei/minimax-h3-opencode-skills

    Default downstream MiniMax H3 specialist for every image-based request unless the user explicitly declares boundary-only first/last frames with no reusable reference role.

    122 GitHub stars~3k tokensUpdated 2 mo ago
    Auto-check passed
  • Minimax H3 Text Video Prompt

    unknowlei/minimax-h3-opencode-skills

    Downstream MiniMax H3 specialist for professional text-to-video (T2VA) prompts using the official three-field format.

    122 GitHub stars~2.3k tokensUpdated 2 mo ago
    Auto-check passed
  • Minimax H3 Creative Director

    unknowlei/minimax-h3-opencode-skills

    Primary mandatory entrypoint for every MiniMax H3 video-generation request.

    122 GitHub stars~2.8k tokensUpdated 2 mo ago
    Auto-check passed
  • Minimax H3 Prompt Reviewer

    unknowlei/minimax-h3-opencode-skills

    Downstream MiniMax H3 specialist that audits, repairs, and rewrites T2VA, I2VA, FL2VA, L2VA, and full-reference prompts into an official structured format.

    122 GitHub stars~2.1k tokensUpdated 2 mo ago
    Auto-check passed

Works with

Questions about Minimax H3 Keyframe Video Prompt

What does Minimax H3 Keyframe Video Prompt do?

Narrow downstream MiniMax H3 specialist for pure I2VA, FL2VA, and L2VA boundary-frame prompts. Minimax H3 Keyframe Video Prompt is an agent skill from unknowlei/minimax-h3-opencode-skills. Narrow downstream MiniMax H3 specialist for pure I2VA, FL2VA, and L2VA boundary-frame prompts.

When should I use Minimax H3 Keyframe Video Prompt?

Minimax H3 Keyframe Video Prompt fits situations like: media & Creative work in your project.

How do I install Minimax H3 Keyframe Video Prompt in Claude Code?

Run `npx skills add unknowlei/minimax-h3-opencode-skills --skill minimax-h3-keyframe-video-prompt -a claude-code`. Or copy the skill folder (skills/minimax-h3-keyframe-video-prompt in unknowlei/minimax-h3-opencode-skills) into .claude/skills/minimax-h3-keyframe-video-prompt in your project. Claude Code loads it when a task matches its description.

How do I install Minimax H3 Keyframe Video Prompt in Codex?

Run `npx skills add unknowlei/minimax-h3-opencode-skills --skill minimax-h3-keyframe-video-prompt -a codex`. Or copy the skill folder (skills/minimax-h3-keyframe-video-prompt in unknowlei/minimax-h3-opencode-skills) into .agents/skills/minimax-h3-keyframe-video-prompt in your project. Codex loads it when a task matches its description.

Can I use Minimax H3 Keyframe Video Prompt in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add unknowlei/minimax-h3-opencode-skills --skill minimax-h3-keyframe-video-prompt -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/minimax-h3-keyframe-video-prompt, .gemini/skills/minimax-h3-keyframe-video-prompt, .github/skills/minimax-h3-keyframe-video-prompt and .opencode/skills/minimax-h3-keyframe-video-prompt in your project.

What does Minimax H3 Keyframe Video Prompt need to run?

SKILL.md names no scripts, command-line tools or credentials: Minimax H3 Keyframe Video Prompt is instructions for the agent only.

Does Minimax H3 Keyframe Video Prompt access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is Minimax H3 Keyframe Video Prompt safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Minimax H3 Keyframe Video Prompt use?

Minimax H3 Keyframe Video Prompt is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Minimax H3 Keyframe Video Prompt use?

About 2.4k tokens (SKILL.md is roughly 9.5k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full. Its references folder adds about 853 tokens, read only when the agent opens those files.

What are the alternatives to Minimax H3 Keyframe Video Prompt?

Skills that share tags, products or a category with Minimax H3 Keyframe Video Prompt: AI Image Generation and Editing (zhayujie/CowAgent, 47k stars), Gif Sticker Maker (HHU3637kr/skills, 145 stars), Minimax H3 Video Reverse (gnipbao/minimax-h3-video-reverse-skill, 230 stars) and H3 Cinematic Director (jtydhr88/ComfyTV, 1.1k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Minimax H3 Keyframe Video Prompt?

unknowlei (a GitHub user) maintains it in unknowlei/minimax-h3-opencode-skills, which has 122 GitHub stars. The repository holds 6 skills in this directory. The repository was last updated on August 9, 2026.

Source: unknowlei/minimax-h3-opencode-skills on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.