Camera & Motion
Direct every camera move
Describe a tracking shot, pan, push-in, or fast follow. H3 Max keeps camera and subject movement aligned with the shot you wrote.
Create short drama shots from text, a first-frame image or reference assets with MiniMax H3 Max on Aniv. Generate 5–15 second videos in 480P, 768P or 1080P with native audio.
Text, keyframes, references and sound — one video workflow
H3 is MiniMax's video model. H3 Max is a post-trained variant tuned for stronger prompt adherence, audiovisual quality and aesthetics. It creates synchronized audio with the picture and follows a written shot more closely than stock H3 at the same resolution, which is why it is the default here.
Direct the shot, keep the look consistent, and bring motion and sound together in one generation.
Camera & Motion
Describe a tracking shot, pan, push-in, or fast follow. H3 Max keeps camera and subject movement aligned with the shot you wrote.
First & Last Frame
Start from an opening image and add an optional closing frame. H3 Max builds continuous motion between the two moments.
Character Consistency
Faces, clothing, proportions, and identity can stay coherent as the setting, light, and camera angle change.
Native Audio
Create ambience, effects, dialogue, and music in the same pass, timed to what happens on screen.
Art Direction
Carry a chosen palette, texture, linework, or cinematic treatment across every beat of the sequence.
Prompt Adherence
Describe the action step by step. H3 Max can follow the sequence and preserve requested visual details, including on-screen text when applicable.
H3 Max turns ideas into motion in seconds. Model inference for a 5-second 768P video completes in about 3 seconds, with queue and file-processing time varying.
Duration, resolution, aspect ratio, seed and end frame — the same fields the model takes.
See the exact cost before generating. Use a subscription or top up when you need more; failed renders return their credits.
H3 reads both, so you can write in either language.
H3 Max prioritizes prompt adherence, audiovisual quality, aesthetics and fast iteration. Standard H3 raises the output ceiling to 2K and 4K. Aniv currently offers H3 Max only.
| Model | H3 MaxDefault here | H3 |
|---|---|---|
| Best for | Prompt fidelity and rapid iteration | 2K and 4K delivery |
| Maximum output | 1080P | 4K |
| Inputs | Text or image, with an optional end frame | Text or image, with an optional end frame |
| Synchronized audio | Included | Included |
Have another question? Contact us at support@aniv.ai
Get started
From a text prompt or an opening frame — a 5 to 15 second cut with synchronized audio, in a single pass.