5. Lock identity by listing the details. "Preserve the half-up long black
hair, openwork silver crown, indigo ribbon, layered pale hanfu, translucent
blue outer robe, deep-blue sash, silver floral fastener, and long tassels."
Naming features gives the model something to hold. The same works for products,
sets and typography.
6. For edits, name the change and the constraint together. Write
substitutions as a list: "Replace the newspaper with a green hardcover book;
replace the chair with a red sofa; remove the subject's sunglasses and reveal a
clear face." Pair each change with what must stay stable and you get a
localized edit instead of a regenerated shot.
7. Use camera and film language. Lens choice, movement, exposure behaviour
and stock character all translate: "subtle handheld shake, then push in quickly
and rack focus", "wide angle with strong perspective distortion", "fine grain,
soft highlight halation, restrained colour", "backlit exposure breathing,
slightly coarse noise in the shadows".
8. Describe transitions as events, not names. "Fast binocular-scan
transitions with whip movement, motion blur, optical smearing, and brief
exposure flicker. Cut at peak blur, then settle and snap back into focus."
Circular vinyl-record wipes, vertical car-door cuts and oversized letter masks
all land better described than labelled.
What it is unusually good at
Reading many references at once, rendering legible text and interfaces, and
making precise localized edits to video you already have. That combination
covers brand films and trailers, title sequences and motion-graphics collage,
vertical short drama, product and e-commerce spots, game and web UI motion,
character and motion transfer, voice cloning, and green-screen replacement — all
from the same model.
Two shapes worth stealing. For a title sequence: name the design language, the
transition vocabulary, the credit typography rules ("each role and each name
appears once"), and a timed BGM brief. For a fashion or product film: assign
mood, talent, product and brand mark to separate images, then keep the story
simple and let the references carry the look.
Check the render with analyze_video, detect_video_scenes and
understand_video; analyze_audio for the mix you directed.
Adapted from fal's MiniMax H3 prompting guide:
https://fal.ai/learn/devs/minimax-h3-prompting-guide