Official guides
MiniMax publishes two official prompt writing guides:- Base generation modes (T2VA, I2VA, FL2VA, L2VA): structure a prompt into timed shots with camera movement and audio (dialogue, SFX, music), with examples for each mode
- Full-reference mode (R2V): the rewrite output structure, including subject definitions, reference labels, and retention analysis, and how to assign each reference a role in the target shot
General tips
- Describe the whole scene: State the overall scene first (location, character, what is happening), then break it into timed shots
- Shots, camera, and audio: Describe the shots, camera moves, and the accompanying audio (dialogue, SFX, music) in one prompt block
- Resolution: H3’s native canvas is a 768px short edge, which is 1344x768 at 16:9, and resolutions are rounded to a multiple of 32. See Setting the output resolution
- Duration: The duration input snaps to the model’s 17-frame-per-block (17k+5) grid at 24fps
Prompt embeddings
Comfy-Org/ComfyUI#15697 added support for prompt embeddings in MiniMax H3. You can use ComfyUI’s standardembedding: syntax in H3 prompts.
Place an embedding file in ComfyUI/models/embeddings/ and reference it in the prompt by name, for example embedding:my_embedding. The embedding is loaded and mixed into the text conditioning just like with any other ComfyUI model.
The Comfy-Org/MiniMax-H3 repository hosts 10 style embeddings in its embeddings folder. These files are unofficial: they were contributed by community member silveroxides via Hugging Face PR #50, and were not produced by Comfy-Org or MiniMax. The original files are in the silveroxides/MiniMax-H3_tests repository. The file name describes the intended effect, and the trigger word is the file name without the extension: