Agent skill

Minimax H3 Prompting

by nodetool-ai in nodetool-ai/nodetool

Prompt MiniMax H3 — choosing between its text-to-video, first-and-last-frame and reference-to-video endpoints, assigning an explicit job to every reference image, clip and audio file, writing a…

AGPL-3.0Auto-check passedMedia & Creative

Install Minimax H3 Prompting

skills CLI
$ npx skills add nodetool-ai/nodetool --skill minimax-h3-prompting -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install nodetool-ai/nodetool minimax-h3-prompting --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/nodetool-ai/nodetool.git skills-src && mkdir -p .claude/skills && cp -r skills-src/packages/system-skills/minimax-h3-prompting .claude/skills/minimax-h3-prompting && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
minimax-h3-prompting
GitHub stars
560
Token cost
~1.6k tokens
SKILL.md length
887 words
Files
1
Skills in repo
127
Repo updated
First seen
Licence
AGPL-3.0

At a glance

Prompt MiniMax H3 — choosing between its text-to-video, first-and-last-frame and reference-to-video endpoints, assigning an explicit job to every reference image, clip and audio file, writing a…

  • The model id starts with minimax/h3 (minimax/h3/text-to-video
  • SKILL.md covers Which endpoint, The eight techniques and What it is unusually good at
  • Instructions only: no scripts, shell commands, URLs or credentials in SKILL.md
  • Minimax/h3/image-to-video

What it does

Minimax H3 Prompting is an agent skill from nodetool-ai/nodetool. Prompt MiniMax H3 — choosing between its text-to-video, first-and-last-frame and reference-to-video endpoints, assigning an explicit job to every reference image, clip and audio file, writing a timecoded shot list inside the prompt, art-directing the native stereo audio, and naming a change and its constraint together when editing a clip. Use whenever the model id starts with minimax/h3 (minimax/h3/text-to-video, minimax/h3/image-to-video, minimax/h3/reference-to-video, their /lora variants, minimax/h3-max/ and…

Its SKILL.md is about 1.6k tokens, which your agent loads only when the skill is triggered. It is a single SKILL.md file with no bundled scripts.

It sits in Media & Creative, covering AI video generation. It works with MiniMax. The repository describes itself as: Agent-first Creative Workspace. The licence is AGPL-3.0.

When your agent uses it

  • The model id starts with minimax/h3 (minimax/h3/text-to-video
  • Minimax/h3/image-to-video
  • Minimax/h3/reference-to-video
  • Their /lora variants

Example prompts

  • “/minimax-h3-prompting”

What it can do on your machine

Read from SKILL.md and the folder at commit 339f069. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    No scripts in the folder and no shell commands in SKILL.md.

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    Links to these hosts (documentation or services it may open):

    • fal.ai

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Minimax H3 Prompting loads about 1.6k tokens when it runs. Until then it costs about 161 tokens; SKILL.md has 887 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~161
When it runs · the whole SKILL.md, loaded when a task matches
~1.6k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from nodetool-ai/nodetool at commit 339f069, republished under its AGPL-3.0 licence (© nodetool-ai). 887 words, ~1,581 tokens.

Download SKILL.mdSave it as .claude/skills/minimax-h3-prompting/SKILL.md (or your agent's skills folder).
name
minimax-h3-prompting
description
Prompt MiniMax H3 — choosing between its text-to-video, first-and-last-frame and reference-to-video endpoints, assigning an explicit job to every reference image, clip and audio file, writing a timecoded shot list inside the prompt, art-directing the native stereo audio, and naming a change and its constraint together when editing a clip. Use whenever the model id starts with minimax/h3 (minimax/h3/text-to-video, minimax/h3/image-to-video, minimax/h3/reference-to-video, their /lora variants, minimax/h3-max/* and minimax/h3-max-turbo/*) on generate_video, animate_image or a TextToVideo node. Not for MiniMax Hailuo.

MiniMax H3 → one context, many references

H3 is a general-purpose multimodal video model rather than a set of task models. Text, images, video and audio all enter one context, so a single request can carry a character's identity from a photo, camera language and cutting rhythm from a clip, and a voice from a recording, and resolve them into one shot with native stereo audio. Output runs 5 to 15 seconds at 24 fps in 2K, and prompts run up to 7,000 characters — a full shot list with sound design fits in one request.

Reach it with find_model for text_to_video or image_to_video, then generate_video / animate_image.

A timecoded shot list in one request is also what keeps the native audio continuous across the cuts — video-audio-continuity for when to fit a whole multi-scene piece into a single generation.

Which endpoint

The rule is mechanical:

  • No media in the request → text to video. Best for anything you can fully describe, and for letting the model invent the look rather than match one. Distinct hand-drawn styles carry surprisingly far on prompt alone.
  • An image that is literally the first or last frame → first and last frame. One image for the opening, or two for opening and closing, with H3 generating the motion between. This is the one for animating a static poster, a UI mockup or key art. Output follows the uploaded image's aspect ratio.
  • Anything you are treating as a reference → reference to video. Up to 9 images, 3 video clips of 2 to 15 seconds, and 3 audio clips, 12 files maximum. Identity locking, motion and camera transfer, style matching, voice cloning, and editing an existing clip all live here. Default to it whenever you have an asset in hand.

The eight techniques

1. Assign a job to every reference. This is the one habit that changes the most. "Use Image 1 for the overall mood, location and film texture; Image 2 for the talent; Image 3 for the bag; and Image 4 for the closing brand mark" beats four images and a description. It works across modalities too: "Match the camera move in Video 1. Make the subject in Video 2 sing, using Video 3 as the reference for both the vocal performance and physical delivery."

2. Write a timed shot list for anything past one beat. Timecoded blocks hold the pacing: "[0 to 2 seconds] High-angle overhead shot… [2 to 4 seconds] Smoothly push in to her right arm… [10 to 15 seconds] As she stands, the full world loads around her." Without them a 15-second generation drifts into a slideshow.

3. Direct the audio. It is generated natively, so art-direct it like a shot: "a deep sub-bass pulse, distant metallic resonance, and one restrained hit as the title locks into focus", or "ice tapping crystal, a faint cigar burn, subtle room air, clothing movement, controlled breathing". For music, describe instrumentation and structure over time, including where the beat lands.

4. State what you do not want. Negative direction is unusually effective and rewards being specific. "No soft dissolves or fluid morphs." "Do not introduce garbled characters or misspellings." "No tearing, black frames, obvious VFX, or compositing seams." These keep a stylized prompt from sliding into a neighbouring genre.

Show full SKILL.md (348 more words)Show less

5. Lock identity by listing the details. "Preserve the half-up long black hair, openwork silver crown, indigo ribbon, layered pale hanfu, translucent blue outer robe, deep-blue sash, silver floral fastener, and long tassels." Naming features gives the model something to hold. The same works for products, sets and typography.

6. For edits, name the change and the constraint together. Write substitutions as a list: "Replace the newspaper with a green hardcover book; replace the chair with a red sofa; remove the subject's sunglasses and reveal a clear face." Pair each change with what must stay stable and you get a localized edit instead of a regenerated shot.

7. Use camera and film language. Lens choice, movement, exposure behaviour and stock character all translate: "subtle handheld shake, then push in quickly and rack focus", "wide angle with strong perspective distortion", "fine grain, soft highlight halation, restrained colour", "backlit exposure breathing, slightly coarse noise in the shadows".

8. Describe transitions as events, not names. "Fast binocular-scan transitions with whip movement, motion blur, optical smearing, and brief exposure flicker. Cut at peak blur, then settle and snap back into focus." Circular vinyl-record wipes, vertical car-door cuts and oversized letter masks all land better described than labelled.

What it is unusually good at

Reading many references at once, rendering legible text and interfaces, and making precise localized edits to video you already have. That combination covers brand films and trailers, title sequences and motion-graphics collage, vertical short drama, product and e-commerce spots, game and web UI motion, character and motion transfer, voice cloning, and green-screen replacement — all from the same model.

Two shapes worth stealing. For a title sequence: name the design language, the transition vocabulary, the credit typography rules ("each role and each name appears once"), and a timed BGM brief. For a fashion or product film: assign mood, talent, product and brand mark to separate images, then keep the story simple and let the references carry the look.

Check the render with analyze_video, detect_video_scenes and understand_video; analyze_audio for the mix you directed.

Adapted from fal's MiniMax H3 prompting guide: https://fal.ai/learn/devs/minimax-h3-prompting-guide

© nodetool-ai, AGPL-3.0. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

Just SKILL.md in packages/system-skills/minimax-h3-prompting of nodetool-ai/nodetool.

Open the folder on GitHubat commit 339f069

Compare with similar skills

Minimax H3 Prompting next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Minimax H3 Prompting compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Minimax H3 Prompting this skillnodetool-ai/nodetool560—~1.6kAutomated safety check: PassAGPL-3.0
Minimax H3 Video Reversegnipbao/minimax-h3-video-reverse-skill230—~2.2kAutomated safety check: PassMIT
Minimax H3 Colabkillkli/minimax-h3-colab-skill110—~743Automated safety check: PassNone
H3 Prompt Writingvllm-project/vllm-omni7.1k6 repos~744Automated safety check: PassApache-2.0
Handdrawn Live Video Generatortl2012tl/comfyUI-llama-TE2414 repos~1.5kAutomated safety check: PassNone
ComfyUI Local DriverSlavaSexton/ComfyUI-Agent-Kit105—~12kAutomated safety check: PassApache-2.0

Similar skills

  • Minimax H3 Video Reverse

    gnipbao/minimax-h3-video-reverse-skill

    先复述参考视频内容与叙事逻辑,和用户确认理解后,再反推 MiniMax H3 文生视频或逐镜图生视频提示词。Use when the user asks to reverse-engineer a video into prompts, reconstruct shots, extract first/end-frame prompts, adapt a reference to H3, or…

    230 GitHub stars~2.2k tokensUpdated 9 days ago
    Media & CreativeAuto-check passed
  • Minimax H3 Colab

    killkli/minimax-h3-colab-skill

    Create short MiniMax H3 reference-to-video clips from one or more local images through Google Colab CLI.

    110 GitHub stars~743 tokensUpdated 15 days ago
    Media & CreativeAuto-check passed
  • H3 Prompt Writing

    vllm-project/vllm-omni

    Write MiniMax H3 video generation prompts for T2VA, I2VA, FL2VA, L2VA, and Ref2VA.

    7.1k GitHub starsUsed in 6 repos~744 tokens
    Media & CreativeAuto-check passed
  • Handdrawn Live Video Generator

    tl2012tl/comfyUI-llama-TE

    For creators making surreal short videos that blend rough glowing hand-drawn animation with live-action spaces.

    241 GitHub starsUsed in 4 repos~1.5k tokens
    Media & CreativeAuto-check passed
  • ComfyUI Local Driver

    SlavaSexton/ComfyUI-Agent-Kit

    Drives a local ComfyUI install over its HTTP API to generate and edit images, video and audio, with per-model prompt recipes and workflow guidance.

    105 GitHub stars~12k tokensUpdated 1 mo ago
    Media & CreativeAuto-check passed
  • Minimax H3 Text Video Prompt

    unknowlei/minimax-h3-opencode-skills

    Downstream MiniMax H3 specialist for professional text-to-video (T2VA) prompts using the official three-field format.

    122 GitHub stars~2.3k tokensUpdated 2 mo ago
    Media & CreativeAuto-check passed

More from nodetool-ai/nodetool

All 127 skills in this repo
  • Beat Sync Editing

    nodetool-ai/nodetool

    Cut a NodeTool timeline to music and shape its pacing — detect the beat grid, place cuts on phrases, pick a cut type, build speed ramps with time remap, and give the piece an arc.

    560 GitHub stars~2.6k tokensUpdated today
    Auto-check passed
  • Caption Titles

    nodetool-ai/nodetool

    Add and animate a consistent text layer on an existing NodeTool timeline.

    560 GitHub stars~1.9k tokensUpdated today
    Auto-check passed
  • Color Motion

    nodetool-ai/nodetool

    Choose and animate colour on a NodeTool timeline, including shape and text gradients, colour grades, 3D LUTs, and dither.

    560 GitHub stars~2.3k tokensUpdated today
    Auto-check passed
  • Commercial Beat Sheet

    nodetool-ai/nodetool

    Write a shootable, precisely timed commercial beat sheet and store it as a NodeTool storyboard, with a consistent entity roster behind every shot.

    560 GitHub stars~4.6k tokensUpdated today
    Auto-check passed
  • Elevenlabs Audio Prompting

    nodetool-ai/nodetool

    Direct ElevenLabs speech, dialogue, sound effects and music — the bracketed audio tags v3 acts on and why the voice decides whether a tag lands, stability as the delivery dial, punctuation instead…

    560 GitHub stars~1.9k tokensUpdated today
    Auto-check passed
  • Frame Composition

    nodetool-ai/nodetool

    Stage the frame on a NodeTool timeline — grids, focal placement, safe areas per aspect ratio, depth layers and parallax, camera moves, and where elements enter and leave.

    560 GitHub stars~3.5k tokensUpdated today
    Auto-check passed

Works with

Questions about Minimax H3 Prompting

What does Minimax H3 Prompting do?

Prompt MiniMax H3 — choosing between its text-to-video, first-and-last-frame and reference-to-video endpoints, assigning an explicit job to every reference image, clip and audio file, writing a…. Minimax H3 Prompting is an agent skill from nodetool-ai/nodetool. Prompt MiniMax H3 — choosing between its text-to-video, first-and-last-frame and reference-to-video endpoints, assigning an explicit job to every reference image, clip and audio file, writing a timecoded shot list inside the prompt, art-directing the native stereo audio, and naming a change and its constraint together when editing a clip.

When should I use Minimax H3 Prompting?

Minimax H3 Prompting fits situations like: the model id starts with minimax/h3 (minimax/h3/text-to-video; minimax/h3/image-to-video; minimax/h3/reference-to-video; their /lora variants.

How do I install Minimax H3 Prompting in Claude Code?

Run `npx skills add nodetool-ai/nodetool --skill minimax-h3-prompting -a claude-code`. Or copy the skill folder (packages/system-skills/minimax-h3-prompting in nodetool-ai/nodetool) into .claude/skills/minimax-h3-prompting in your project. Claude Code loads it when a task matches its description.

How do I install Minimax H3 Prompting in Codex?

Run `npx skills add nodetool-ai/nodetool --skill minimax-h3-prompting -a codex`. Or copy the skill folder (packages/system-skills/minimax-h3-prompting in nodetool-ai/nodetool) into .agents/skills/minimax-h3-prompting in your project. Codex loads it when a task matches its description.

Can I use Minimax H3 Prompting in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add nodetool-ai/nodetool --skill minimax-h3-prompting -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/minimax-h3-prompting, .gemini/skills/minimax-h3-prompting, .github/skills/minimax-h3-prompting and .opencode/skills/minimax-h3-prompting in your project.

What does Minimax H3 Prompting need to run?

SKILL.md names no scripts, command-line tools or credentials: Minimax H3 Prompting is instructions for the agent only.

Does Minimax H3 Prompting access the network?

SKILL.md names 1 domain. As links in the text: fal.ai. This is read from the text; nothing was executed.

Is Minimax H3 Prompting safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Minimax H3 Prompting use?

Minimax H3 Prompting is published under the AGPL-3.0 licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Minimax H3 Prompting use?

About 1.6k tokens (SKILL.md is roughly 6.3k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Minimax H3 Prompting?

Skills that share tags, products or a category with Minimax H3 Prompting: Minimax H3 Video Reverse (gnipbao/minimax-h3-video-reverse-skill, 230 stars), Minimax H3 Colab (killkli/minimax-h3-colab-skill, 110 stars), H3 Prompt Writing (vllm-project/vllm-omni, 7.1k stars) and Handdrawn Live Video Generator (tl2012tl/comfyUI-llama-TE, 241 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Minimax H3 Prompting?

nodetool-ai (a GitHub organization) maintains it in nodetool-ai/nodetool, which has 560 GitHub stars. The repository holds 127 skills in this directory. The repository was last updated on October 10, 2026.

Source: nodetool-ai/nodetool on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.