Hailuo AI / MiniMax 2.3
MiniMax's Hailuo 2.3 delivers high-motion, physics-aware video clips with subject reference and flat per-clip billing.

Picking Hailuo 2.3 as the model for a generation.
What Hailuo 2.3 does
Hailuo 2.3 converts a text prompt or reference image into a short video clip, with particular strength in scenes involving movement, action, or physical interaction — a subject walking, an object being thrown, fabric or hair reacting to motion. It's positioned as a workhorse model for shots where believable physics matters more than stylistic flourish.
Its first-and-last-frame mode and subject reference inputs make it well suited to shots that need to hit a specific start and end state, or that need to keep a particular face or character consistent across a sequence of separate generations.
Key features
High-motion, physics-aware generation
Hailuo 2.3 is specifically tuned for realistic movement and physical interaction — collisions, falls, fabric and hair dynamics — rather than generic scene composition, making it a strong choice for action-heavy shots.
First-and-last-frame control
Supply a starting frame and an ending frame and the model interpolates the motion between them, giving precise control over how a shot opens and resolves rather than leaving the endpoint to chance.
Subject reference (face and body)
Reference a specific person's face and body across separate generations to keep a character recognizable through a multi-shot sequence without regenerating their appearance from scratch each time.
Flat per-clip billing
Cost is billed at a flat rate per clip based on the maximum clip length, so the price is known before you generate rather than fluctuating with the actual rendered duration.
Fast 6-10 second turnaround
Shorter default clip lengths keep iteration fast when you're testing prompt variations or blocking out a sequence before committing to longer chained shots.
How Hailuo 2.3 works
Hailuo 2.3 is a diffusion-based video model tuned specifically for motion realism and physical plausibility — how objects fall, collide, deform, and interact with each other on screen. MiniMax has iterated the Hailuo line release-over-release specifically on this axis, and 2.3 continues that focus with visible improvements in how bodies move and how physical forces play out across a shot.
It accepts both text prompts and reference images as inputs, and supports first-and-last-frame control — supplying a starting frame and an ending frame so the model interpolates the motion between them, which is useful for locking down exactly how a shot should begin and resolve. Subject reference conditioning (face and body) lets a specific person or character stay recognizable across separate generations.
Clips run 6-10 seconds at 720p-1080p by default, with select configurations available up to 4K. Billing on AutorunX is flat per clip at the maximum clip length rather than scaling with actual output length, so cost is predictable up front regardless of how long the final rendered clip turns out to be.
What people use Hailuo 2.3 for
Action and movement shots
Generate clips where physical motion is the focus — a product being handled, a subject walking or turning, fabric or hair responding to movement.
Locked start/end shots
Use first-and-last-frame control when a shot needs to open and close on specific compositions, such as a product reveal that must end on a clean hero frame.
Recurring character shots
Use subject reference to keep the same face and body consistent across multiple separate Hailuo generations in a sequence.
Fast prompt iteration
Lean on the shorter 6-10 second default clips to test several prompt or reference variations quickly before committing to a final take.
Who built Hailuo 2.3
MiniMax
www.minimax.ioMiniMax is a Chinese AI company building large multimodal foundation models, including the Hailuo video generation line. It has iterated quickly through the Hailuo series, positioning successive releases around motion realism, physical plausibility, and production-ready output for creative and marketing teams.
How to use Hailuo 2.3 on AutorunX
Hailuo 2.3 is available in Short Film in Video Lab.
Open Short Film or Explainer in Video Lab
From the AutorunX dashboard, go to Video Lab and open either Short Film or Explainer — Hailuo 2.3 is available as a model choice in both.
Set up your input
Write a prompt, attach a reference image for subject consistency, or supply a first and last frame if you need a locked start and end composition.
Pick Hailuo 2.3 in the model picker
Open the model picker and select Hailuo 2.3 — it's a selectable option, so confirm it's chosen before generating.
Generate and review
Hit generate to render a 6-10 second clip. Review the motion and physical interaction, then chain or regenerate as needed.

An example generation rendered with Hailuo 2.3.
Credit usage
Billed per clip from your shared AutorunX credit wallet.
Tips for better results with Hailuo 2.3
Describe physical forces explicitly
Naming what should physically happen — a fall, a gust of wind, an impact — gives Hailuo a clearer target than a purely visual description of the end state.
Use first-and-last-frame for precise pacing
When a shot needs to land on an exact final composition, supply both endpoints instead of relying on the model to guess where the motion should stop.
Reference the same subject image across a sequence
Reuse the same face/body reference image across multiple generations rather than a new one each time to reduce drift in a character's appearance.
Budget for the flat rate
Since billing is flat per clip at max length, there's no cost benefit to requesting a shorter clip — plan the shot length around the content, not the price.
Hailuo 2.3 — frequently asked questions
Related models
Seedance 2.0
ByteDance's default AutorunX video engine for Short Film, Movie Maker, and Ad Remake — identity-preserving reference-to-video at up to 4K.
Video GenerationKling 3.0
Kuaishou's flagship video model with native multi-lingual audio, in-video editing, and clips up to native 4K.
Video GenerationVeo 3.1
Google's flagship text-to-video and image-to-video model with native synced audio.
Ready to create with Hailuo 2.3?
Sign up for AutorunX to get 200 free credits across every lab, including Short Film.