Model reference · Generator

MiniMax H3 Max

MiniMax H3 Max is a post-trained version of MiniMax H3 developed by fal Research, rather than a separate “Max” edition released by MiniMax. fal starts from the open-weight H3 base, adds post-training focused on prompt adherence and visual quality, and serves the result through an inference stack co-optimized for high throughput.

The current H3 Max endpoints focus on fast text-to-video and image-to-video generation at 480P or 768P with synchronized audio. Standard MiniMax H3 remains the better fit when you need 2K output or its broader reference-to-video workflows.

H3 Max by fal Research

Generate with MiniMax H3 Max

Create a video from text or animate a first frame with synchronized audio.

H3 Max can expand a simple prompt before generation.0/7,000
Resolution

Need 2K output? MiniMax H3 supports 2K workflows.

Duration
5s
Aspect Ratio
Prompt Expansion

Improve prompt structure with minimal added latency.

Advanced Settings

Key facts

MiniMax H3 Max at a Glance

MiniMax H3 Max key facts
AttributeCurrent fact
Developerfal Research
Base modelMiniMax H3
ReleaseAugust 27, 2026
TypePost-trained H3 variant
Public launch workflowsText to Video / Image to Video
Resolution480P / 768P
Default resolution768P
Duration5–15 seconds
Text-to-video aspect ratios21:9 / 16:9 / 4:3 / 1:1 / 3:4 / 9:16
Image-to-video compositionFollows first-frame aspect ratio
Last-frame controlOptional end image
AudioSynchronized/native audiovisual generation
Prompt expansionDisabled / Balanced / Quality
Primary positioningSpeed + prompt adherence + aesthetics

What Is MiniMax H3 Max?

MiniMax H3 Max is fal Research's post-trained variant of the open-weight MiniMax H3 model. fal added substantial post-training data with an emphasis on prompt adherence and visual quality, while its inference team co-optimized the serving stack for much higher throughput.

H3 Max should therefore be understood as a derived H3 model and serving system, not simply as a higher-resolution or premium-named version of MiniMax H3.

Read fal's H3 Max launch article

What Actually Makes H3 Max Different?

Model Post-Training

fal starts with MiniMax H3 open weights, adds substantial new training data, and evaluates checkpoints for prompt understanding and aesthetics. Its launch materials also describe reinforcement-learning work with verifiable rewards.

Inference-System Optimization

fal developed the serving stack alongside the model, applying custom diffusion inference and kernel work while preserving the post-training quality gains. Published throughput therefore reflects the combined model and system.

H3 Max's speed is not only a model-weight story. fal describes H3 Max as the result of post-training and inference-system co-design, so its published latency reflects both the derived model and the serving stack.

Comparison

MiniMax H3 Max vs MiniMax H3

H3 Max compared with MiniMax H3
AttributeH3 MaxMiniMax H3
Developerfal ResearchMiniMax
RelationshipPost-trained H3 derivativeBase/original model family
Primary focusSpeed + prompt adherence + aestheticsBroader general-purpose multimodal workflows
Resolution480P / 768PUp to 2K
Duration5–15 sec5–15 sec on current fal H3 endpoints
Native/synchronized audioYesYes
Text to VideoYesYes
Image to VideoYesYes
First + Last FrameYesYes
Reference to VideoNo live public Max endpoint verifiedYes
Rich image/video/audio referencesNo equivalent current public Max workflow verifiedYes
Separate public H3 Max weightsNo separate release currently statedH3 open weights available
Best fitFast iteration / throughput2K / richer multimodal reference control

H3 Max is not simply H3 with higher specifications. Its main advantages are fal's post-training and serving speed, while standard MiniMax H3 retains 2K output and broader reference-to-video workflows.

Should You Use H3 Max or MiniMax H3?

Choose H3 Max when

  • generation latency matters
  • you iterate prompts frequently
  • 768P is sufficient
  • your workflow is primarily T2V or I2V
  • prompt adherence is a priority

Choose MiniMax H3 when

  • you need 2K
  • you need richer reference-to-video inputs
  • you need broader image/video/audio reference workflows
  • you specifically need the standard/open H3 ecosystem

H3 Max is primarily a speed-and-adherence option, not a universal replacement for MiniMax H3.

Decision map showing H3 Max for fast 768P iteration and MiniMax H3 for 2K or richer multimodal references
When to use fal's H3 Max versus standard MiniMax H3.

Evidence

How Fast Is MiniMax H3 Max?

fal reports that H3 Max can generate a 5-second 768P video in under three seconds on its optimized serving stack. fal also reports roughly 35× the throughput of the official H3 endpoint in its launch comparison.

Vendor-reported launch benchmark

H3 Max speed and ranking evidence
ClaimEvidenceType
5-second clip in under ~3 secondsfal launch benchmarkVendor-reported
~35× official H3 endpoint throughputfal launch benchmarkVendor-reported
Design Arena rankingExternal leaderboard cited by falLeaderboard snapshot
Artificial Analysis rankingExternal leaderboard cited by falLeaderboard snapshot

Leaderboard positions and provider benchmarks are snapshots, not permanent model specifications.

MiniMax H3 Max Specs

Resolution

480P / 768P

Duration

5–15 seconds

T2V ratios

21:9 · 16:9 · 4:3 · 1:1 · 3:4 · 9:16

I2V input

First frame; optional last frame

I2V composition

Follows first image

Prompt expansion

Disabled · Balanced · Quality

Audio

Synchronized audio

Safety

Provider safety checking stays enabled

H3 Max Prompt Expansion: Disabled vs Balanced vs Quality

H3 Max prompt expansion modes
ModeBehavior
DisabledSend the original prompt without fal's expansion step
BalancedFast prompt expansion; default/recommended
QualitySpend additional time producing a richer expanded prompt

The expanded prompt is a preprocessing layer exposed by fal's H3 Max API. It is not the H3 Max model itself.

Learn the MiniMax H3 prompt structure

MiniMax H3 Max Pricing

Provider price snapshot

fal API pricing

As checked August 29, 2026, fal lists promotional launch rates of $0.025/second at 480P and $0.04/second at 768P. fal states that the promotion ends September 1, after which list rates are $0.05/second and $0.08/second respectively.

Check current provider pricing

Estimated minimax3.org cost

Generation credits

The generator quote uses minimax3.org credits per output second. These retail credits are configured independently from fal's provider pricing and the current estimate shown before generation is authoritative.

480P
4 credits/sec
768P
6 credits/sec

MiniMax H3 Max API

TEXT TO VIDEO

minimax/h3-max/text-to-video

prompt · duration · resolution · prompt_expansion_mode · aspect_ratio · seed

IMAGE TO VIDEO

minimax/h3-max/image-to-video

prompt · duration · resolution · prompt_expansion_mode · image_url · end_image_url · seed

Relevant output fields are video, expanded_prompt, and timings.inference. Provider keys must stay on the server.

Text-to-video API docsImage-to-video API docs

Is MiniMax H3 Max Open Source?

MiniMax H3 itself has open weights. H3 Max is a fal Research post-trained derivative, but fal's current launch materials do not announce a separate downloadable H3 Max checkpoint. Users currently access H3 Max through fal's hosted product and API.

If separate H3 Max weights are released later, this section should be updated.

MiniMax H3 Max FAQ

What is MiniMax H3 Max?

H3 Max is fal Research's post-trained variant of the open-weight MiniMax H3 model, co-optimized with fal's inference stack for speed, prompt adherence and visual quality.

Is H3 Max made by MiniMax?

No. MiniMax developed the H3 base model. fal Research developed the H3 Max post-trained derivative and serving system.

Is H3 Max better than MiniMax H3?

It depends on the workflow. H3 Max prioritizes speed and prompt adherence at up to 768P; standard H3 is the stronger fit for 2K and broader multimodal reference workflows.

How fast is H3 Max?

fal reports a 5-second video in under roughly three seconds on its optimized serving stack. This is a vendor-reported launch benchmark, not a guaranteed end-to-end request time.

What resolution does H3 Max support?

The current public H3 Max endpoints expose 480P and 768P output.

How long can H3 Max videos be?

Current fal product documentation supports single generations from 5 to 15 seconds.

Does H3 Max generate audio?

Yes. fal describes H3 Max as retaining H3’s natively synchronized audio and video capability.

Does H3 Max support image-to-video?

Yes. Its image-to-video endpoint accepts a first-frame image and derives the output composition from that image.

Can H3 Max use a first and last frame?

Yes. Add an optional end_image_url to the image-to-video endpoint to control the final frame.

What is H3 Max prompt expansion?

It is a fal API preprocessing control that rewrites the prompt before generation. Modes are disabled, balanced and quality.

Does H3 Max support 2K?

The current H3 Max endpoints do not expose 2K. Standard MiniMax H3 supports 2K workflows.

Is MiniMax H3 Max open source?

MiniMax H3 has open weights. fal’s current H3 Max launch materials do not announce a separate downloadable H3 Max checkpoint.

Can I download H3 Max weights?

No separate H3 Max checkpoint is currently announced. H3 Max is currently accessed through fal’s hosted product and API.

What is the H3 Max API?

fal exposes minimax/h3-max/text-to-video and minimax/h3-max/image-to-video through its queued model API.

Source ledger

H3 Max Sources & Verification

H3 Max claims and primary sources
ClaimSource
H3 Max is fal's post-trained H3 derivativefal H3 Max launch article
Post-training + inference co-designfal H3 Max launch article
480P / 768P and T2V ratiosfal T2V API schema
5–15 second supportfal H3 Max product documentation
I2V follows first image aspect ratiofal I2V API schema
Optional final framefal I2V API schema
Prompt expansion, expanded_prompt, timings.inferencefal API output schema
Standard H3 supports 2K / richer referencesfal standard H3 documentation

Last verified: August 29, 2026.

Download the machine-readable comparison dataset