Independent launch guide · Updated July 31, 2026

MiniMax H3,
decoded.

Verified specifications, launch analysis, prompt craft, and availability updates for MiniMax’s new general-purpose multimodal video model.

Verified July 31, 2026. Specifications on this site are checked against MiniMax's release notes and technical documentation. Unknowns stay labeled as unknowns.

2K

maximum output

4–15s

integer duration

12

mixed references

7,000

prompt characters

One model, more context

A multimodal video workflow

MiniMax describes H3 as a general-purpose model that understands text, image, video, and audio inputs in a unified generation workflow.

01

Text to video

Direct a full scene from a text brief, with prompts up to 7,000 characters.

02

Image guidance

Use first and last frames or image references to anchor composition and identity.

03

Multimodal references

Combine up to 12 total reference files, subject to per-media limits.

04

Audio references

Provide audio references alongside image and video inputs for generation direction.

Demo desk

Evidence over a highlight reel

We do not download or re-upload creator work. Until official platform embeds are available, the demo desk links directly to MiniMax’s documented examples and keeps analysis separate from launch claims.

Access status

No fake generator button

The fastest path to trust is saying exactly what works today.

API documented; this site is not the generator

MiniMax documents MiniMax-H3 on its video generation API. MiniMaxH3.xyz currently provides independent guidance and a local prompt builder—it does not upload media or send generation jobs.

Open official API documentation

Build with clarity

Start with a stronger prompt.

Compose scene, camera, lighting, motion, and audio direction in one structured brief.

Open prompt builder
MiniMax H3: Features, Demos & Updates | MiniMaxH3