Example input
News report
Post-trained by fal Research
Turn detailed shot briefs into polished motion faster, with stronger prompt adherence and clear visual direction.
Start from a shot briefPreview
Example input
News report
Loading history…
Enter to generate · Shift + Enter for a new line
0/5000
Story Workshop
Decide how you want AI to help
Choose a conversation starter, review the brief, then open it in your preferred AI chat.
Choose a starting prompt.
The AI will guide the discussion and finish with one copy-ready video prompt.
Open in your preferred AI chat.
Your prepared prompt is sent to the service you choose and may appear in its URL, browser or account history, and logs. That provider's terms and privacy policy apply.
Watch the model work
fal's release gallery shows where its post-training work lands: prompt-led transformations, consistent characters, performance, and kinetic typography.

A staged inflation sequence guided by visual keyframes.
A stylized character shot that holds the subject through motion.
A small ensemble performance with synchronized scene energy.
A typographic transition that tests timing and visual adherence.
Media is sourced from the model publisher's official showcase.
Why this model matters
fal tuned H3 Max for the moment after the idea is clear: execute more of the brief, return iterations faster, and keep the final frame polished.
Spell out ordered actions, camera movement, dialogue, and constraints in one brief with less simplification.
fal reports a five-second 768p clip can complete in under three seconds of backend generation time.
Use a start image and optional end image to direct a product reveal, transformation, or transition.
Keep subject identity, materials, lighting, and composition coherent while the action unfolds.
Start with direction
H3 Max is built for dense, ordered direction. Start with one of these shot briefs, then tune the timing, framing, movement, and look in the generator.
A precise sequence of actions with a clean final beat.
Fast product variations with explicit brand framing.
A text-directed transformation with an exact landing frame.
Gesture, reaction, and camera instructions in one take.
Same foundation, different priorities
Choose H3 Max for faster 480p or 768p iteration and stronger adherence to detailed shot instructions. Choose base H3 for its broader editing story and up-to-2K path.
Best for
MiniMax H3
Omni-modal control and a higher-resolution finish
H3 Max by fal
Fast, prompt-faithful iteration
Resolution
MiniMax H3
Up to 2K through regeneration
H3 Max by fal
480p or 768p
Direction
MiniMax H3
Text, frames, images, video, and audio
H3 Max by fal
Text, start/end images, and references
Core strengths
MiniMax H3
Reference control, native stereo, and editing
H3 Max by fal
Speed, adherence, aesthetics, and visual coherence
Prompt recipe
A strong video brief reads like compact direction, not a bag of style words. Give the model an ordered scene it can execute.
Straight answers
fal Research post-trained H3 Max on the open-weight MiniMax H3 base model and co-optimized it with fal's inference team. MiniMax did not release H3 Max as a separate base model.
fal positions H3 Max around higher throughput, stronger prompt adherence, and improved aesthetics. The base H3 page emphasizes broader multimodal editing and an up-to-2K regeneration workflow.
fal exposes text-to-video, image-to-video with an optional end frame, and reference-to-video endpoints for H3 Max. This page's generator uses the text-to-video route.
fal lists 480p and 768p output for H3 Max, with clips from 5 to 15 seconds at 24 FPS.