Models / Hailuo (MiniMax) / Video to video
Hailuo (MiniMax) video to video prompting guide
Last checked 2026-10-11Hailuo / MiniMax H3 · MiniMax H3 (hailuo3: t2v/i2v/r2v/v2v), H3 Max (h3_max: t2v/i2v first+last), Hailuo 2.3
How to prompt Hailuo (MiniMax) in video-to-video mode. Every point links its source; community tips are marked. Your agent gets the same guide through the Atlas MCP (get_guide).
Prompt structure
video binding→edit or transfer instruction→ref jobs→soundscape
- Three fields: integrated_multimodal_description (visuals, action, shots, dialogue on the timeline), overall_soundscape (1–4 sentences), non_diegetic_music (or N/A). source ↗
What the text should describe
- Refer to uploads by modality and order: Image 1, Video 1, Audio 1. Give each clip a job ('Match the camera move in Video 1', 'Make the subject in Video 2 sing using Video 3 as the vocal reference'). source ↗
- Runway hailuo3 video_to_video: promptVideo + promptText; extra video refs share a 15 s total budget. source ↗
Camera vocabulary
- Camera = motion type + amplitude + speed as a sentence: 'The camera pushes in with small amplitude at slow speed toward ...'. Types: zoom, push/pull, pan, truck, tilt, pedestal, arc, tracking, static, shake, POV, roll. source ↗
Responds well to
- On-screen text in double quotes, verbatim. source ↗
Duration and limits
- H3 Reference to Video: up to 9 images, 3 videos (2–15 s), 3 audio, 12 files; Runway hailuo3 768p/2k, h3_max 480p/768p. source ↗
Multi-shot
- [Shot 1] (no timestamp), then '[Shot 2] At 00:03.500, the camera cuts to ...' with strictly increasing times. source ↗
Avoid and failure modes
- Use a cut only for new information; small angle changes = camera motion. source ↗
Negative prompt
Supported via .
Next
All Hailuo (MiniMax) modesHailuo (MiniMax) model, prices & shotsFrom $0.1/s · pricesSeedance video to videoKling video to videoRunway video to videoLuma video to video