MiniMax H3 Max is now in Picsart AI Playground, returning a 5-second clip in seconds rather than minutes. It generates video from a text prompt or an image at 480p or 768p, in any length from 5 to 15 seconds, and it is the speed-tuned member of the MiniMax H3 family.
Speed sounds like a minor spec until it changes the order you work in. A model that answers quickly turns generation into something you iterate with rather than something you commission and wait on. That shift, not the resolution number, is what makes H3 Max worth reaching for.
Here is what the model gives you, the settings behind it, and a way of working that a fast model makes possible.
What MiniMax H3 Max is
MiniMax H3 Max is a post-trained variant of MiniMax H3, tuned for stronger prompt adherence and better aesthetics, and co-optimized for higher throughput. It is not simply a lower-resolution H3. The only thing it gives up is resolution headroom, and what it gains is turnaround plus a tighter reading of the prompt.
It works two ways:
- Text to video. Write what you want to see and the model builds it with no source material.
- Image to video. Supply a start frame, an end frame, or both, and the model builds the motion between them.
Both routes run through the same prompt box, and supplying no image simply makes it a text-to-video job. 768p is the ceiling here. 480p sits below it, which is plenty while a draft only has to communicate blocking and timing.
The settings behind it
| Setting | What you can choose |
|---|---|
| Resolution | 480p or 768p |
| Aspect ratio | 21:9, 16:9, 4:3, 1:1, 3:4, 9:16 |
| Duration | Any whole number of seconds, 5 to 15 |
| Starting point | Text prompt, a start frame, or a start and end frame |
| Frame rate | 24 fps |
| Prompt expansion | Disabled, balanced, or quality |
| Seed | Optional, for repeatable results |
Duration is a slider, not a set of fixed options. A 7-second clip is as available as a 5 or a 15, which matters when a cut has to land on a specific beat.
Cost scales with length. A 5-second draft is a third of the commitment of a 15-second one, which is what makes exploration affordable.
Prompt expansion has three settings, not a switch. Disabled holds you to exactly what you wrote. Balanced fills in the gaps a short brief leaves. Quality elaborates hardest, which suits a one-line idea you want fleshed out. Moving between them changes the output more than most people expect.
A seed makes a result repeatable. Set one and the same prompt returns the same clip, which is what lets you change a single word and see only that word’s effect.
Why a fast model changes how you work
Generation speed sets the size of the loop you can afford. A slow model pushes you toward writing one careful prompt and accepting what comes back, because a second attempt costs real time. A fast model lets you be wrong cheaply.
That has a practical consequence. Framing, pacing, and camera movement are difficult to predict from a text description and easy to judge on sight. With a fast model you stop predicting and start looking: generate four versions, watch them, and let the one that works tell you what the prompt should have said.
The corollary is that a first generation stops being a deliverable and becomes a question. Most of the value arrives in the third or fourth attempt, and only a model that answers quickly makes getting there realistic.
A drafting workflow that fits it
Fast output rewards a different sequence than careful output does. This one works:
- Draft at 480p and 5 seconds. Cheapest possible look at whether the idea reads at all. Composition and subject are visible; nothing else matters yet.
- Fan out on the version that works. Regenerate the same prompt three or four times to see the range the model gives you. Variation at this stage is information, not noise.
- Fix the framing before the length. Lock the aspect ratio and the composition while clips are still short, because a reframe late costs the edges of the shot.
- Extend to the real duration. Move to the length the cut actually needs, now that the shot is settled.
- Finish at 768p. Raise the resolution last, on the one version that is going out.
The order matters more than the individual steps. Every decision that is cheap to change gets made while generations are still short and low-resolution.
What to make with it
H3 Max suits work where the count is high and the format is fixed:
- Vertical social clips. 9:16 at 768p, sized for where most short video is watched.
- Square feed posts. 1:1 for placements that crop a widescreen clip badly.
- Widescreen and cinematic. 16:9 for site headers and 21:9 when a shot needs to read as film.
- Animated stills. A product photo or a piece of key art given motion, using the start frame.
- Transitions with a fixed destination. Both frames supplied, letting the model solve the move between two images you already have.
Across all of them the pattern is the same: several versions, judged by eye, with the winner taken up to full resolution.
Tips for better results
Write the motion, not just the scene
A still description produces a still-feeling clip. Say what moves and in what direction.
Use an end frame when the shot has a destination
A move that must land somewhere specific is far more reliable with both ends supplied.
Change one thing per attempt
Rewriting the whole prompt between generations makes it impossible to tell what caused the improvement.
Keep drafts at 480p
Resolution is the last thing worth spending on, and the cheapest thing to raise.
Pick the ratio first
It shapes composition, so deciding it late means regenerating rather than cropping.
Where MiniMax H3 Max sits in Picsart
Pick MiniMax H3 Max from the video model dropdown in Picsart AI Playground, then set resolution, ratio, and length from the bar beside it. Nothing needs installing and no parameters need wiring up. It runs on the web and in the Picsart desktop app.
Image and music generation share that same canvas, which keeps a clip and the assets around it inside one project. The Picsart AI models catalog holds more than 150 models in total, MiniMax H3 among them for jobs that call for reference inputs or a higher-resolution finish. For trimming and assembly afterwards, the Picsart AI video generator covers the surrounding workflow.
Get answers to common questions
MiniMax H3 Max is a speed-tuned variant of MiniMax H3, post-trained for stronger prompt adherence and better aesthetics. It generates from a text prompt or an image at 480p or 768p, in lengths from 5 to 15 seconds, and returns a 5-second clip in seconds rather than minutes.
Start generating with MiniMax H3 Max
Open Picsart AI Playground, pick MiniMax H3 Max, and run a 5-second draft at 480p. Generate it four times before you judge it. The version that surprises you is the one worth taking up to 768p.