Verified model snapshot
What changed
Where to access this model
fal
9 observed offers · from $0.025/generated second
Last checked Sep 22, 2026
View 9 offer details
USD 0.025 / generated second
$0.025 per generated second; 50% promotional rate
Regular rate is $0.05 per second at 480P.
Source ↗USD 0.04 / generated second
$0.04 per generated second; 50% promotional rate
Regular rate is $0.08 per second at 768P.
Source ↗USD 0.08 / generated second
$0.08 per generated second; 50% promotional rate
1080P is latent refinement from a native 768P source; regular rate is $0.16 per second.
Source ↗USD 0.08 / generated second
$0.08 per generated output second
Generated-output component only; reference inputs are billed separately by pooled tokens.
Source ↗USD 0.02 / per 1000 reference tokens over 4096 free
$0.02 per 1,000 reference tokens over 4,096 free
Do not normalize with output-video per-second pricing.
Source ↗USD 0.05 / generated second
$0.05 per generated second
Source ↗USD 0.08 / generated second
$0.08 per generated second
Source ↗USD 0.16 / generated second
$0.16 per generated second
Source ↗USD 0.32 / generated second
$0.32 per generated second
Source ↗Pixazo
3 observed offers · from $0.025/generated second
Last checked Sep 22, 2026
View 3 offer details
USD 0.025 / generated second
$0.025 per generated second; current promotional table
Current comparable rate; no unique-cheapest claim is made.
Source ↗USD 0.04 / generated second
$0.04 per generated second; current promotional table
1080P is described as refinement from native 768P.
Source ↗USD 0.08 / generated second
$0.08 per generated second; current promotional table
1080P is described as refinement from native 768P.
Source ↗Runware
2 observed offers · from $0.025/generated second
Last checked Sep 22, 2026
Cloudflare Workers AI
1 observed offer
Last checked Sep 22, 2026
View 1 offer details
Public page routes pricing to the dashboard; no stable raw USD rate captured.
Source ↗Layer
6 observed offers
Last checked Sep 22, 2026
View 6 offer details
0.42 Creative Units per second
Raw Creative Units retained; no USD conversion.
Source ↗0.672 Creative Units per second
Raw Creative Units retained; no USD conversion.
Source ↗1.344 Creative Units per second
Raw Creative Units retained; no USD conversion.
Source ↗1.67 Creative Units per second
Raw Creative Units retained; no USD conversion.
Source ↗2.672 Creative Units per second
Raw Creative Units retained; no USD conversion.
Source ↗5.344 Creative Units per second
Raw Creative Units retained; no USD conversion.
Source ↗What stands out
Evidence-backed signals
- Native audio is positively verified from an entity-supporting source.
- Maximum duration is recorded as 5–15 seconds.
- Official API access is verified as available.
Objective collection placement
Recorded test runs
Independent reports and examples
These observations belong to the named publishers and providers; they are not FrameSignal hands-on tests.
submit to result wait
A 5-second 768P product turntable completed in one attempt. Voyager reports no reroll or retouch in the common nine-clip fixture; the listed cost is an estimate from the published promotional rate.
Reported metric: 6 seconds
Reported result: 5s clip; one attempt; estimated cost $0.20 at the published promotional rate.
Prompt: A clean product shot that holds its shape through a full slow rotation. (Full prompt preserved at source.)
Settings: endpoint: minimax/h3-max/text-to-video · resolution: 768P · duration seconds: 5 · aspect ratio: 16:9 · prompt expansion mode: disabled
Open source ↗submit to result wait
A busy 768P text-to-video scene retained distinct moving people and objects in Voyager's test, while submit-to-result wait was materially longer than the backend inference examples.
Reported metric: 26 seconds
Reported result: Distinct moving people and objects without merging; estimated cost $0.20.
Prompt: A crowded Saturday farmers market seen from slightly above, shoppers with tote bags, a dog on a leash, a busker with a guitar, bunting swaying, late morning light, gentle camera drift.
Settings: endpoint: minimax/h3-max/text-to-video · resolution: 768P · duration seconds: 5 · aspect ratio: 16:9 · prompt expansion mode: disabled
Open source ↗submit to result wait
An image-to-video test at 768P kept the source silhouette recognizable, but Voyager reports that the dripper filled like a cup rather than following the requested physical action.
Reported metric: 6 seconds
Reported result: Output returned at 768x768; source silhouette remained recognizable but the dripper filled like a cup.
Prompt: The coffee dripper slowly fills as coffee drips through it, steam rising, the camera drifting gently around it, the dripper itself unchanged.
Settings: endpoint: minimax/h3-max/image-to-video · resolution: 768P · duration seconds: 5 · prompt expansion mode: disabled
Open source ↗submit to result wait
The camera-control test kept the product recognizable; Voyager cautions that precise orbit accuracy is difficult to establish from the near-symmetric object used.
Reported metric: 6 seconds
Reported result: Product remained recognizable; precise 3D reconstruction was not established.
Prompt: The coffee dripper is rigid and motionless. Only the camera moves. Preserve its shape and materials and the white background.
Settings: endpoint: minimax/h3-max/camera-controls · resolution: 768P · duration seconds: 5 · camera trajectory: [{"time":0,"azimuth":0,"elevation":0,"distance":1},{"time":1,"azimuth":45,"elevation":10,"distance":1}]
Open source ↗submit to result wait
A lip-sync adapter test retained the supplied audio and changed mouth shape. Voyager explicitly did not score synchronization quality by ear.
Reported metric: 11 seconds
Reported result: Estimated cost $0.6895; supplied audio was retained and mouth shape changed.
Prompt: Animate the generated portrait to the supplied synthetic ad-read audio.
Settings: endpoint: minimax/h3-max/lip-sync/image-to-video · resolution: 480P · duration seconds: 14
Open source ↗reported result
In a 768P mechanical-transformation test, the main assembly was readable and the design survived into wing beating, but the requested fly-past was missed by the late frame.
Reported result: Main assembly was readable and design survived into wing beating, but the requested fly-past was missed by the late frame.
Prompt: One continuous 5-second macro shot in a sunlit watchmaker's workshop... loose brass gears and cobalt-blue plates rapidly snap together into one mechanical hummingbird... then launches forward before arcing past the camera. (Full prompt preserved at source.)
Settings: resolution: 768P · duration seconds: 5 · aspect ratio: 16:9 · passes: 1 · prompt expansion mode: balanced · safety checker: true
Open source ↗submit to result wait
A 768P vertical action test measured 6.97 seconds submit-to-result wait while fal inference was reported as 2.79 seconds. Subject and motion remained coherent in sampled frames; fine wheel geometry softened at peak speed.
Reported metric: 6.97 seconds
Reported result: fal inference 2.79s; measured wait 6.97s. Subject and motion remained coherent; fine wheel geometry softened at peak speed.
Prompt: One continuous 5-second vertical documentary shot on a rainy city side street at blue hour... bicycle messenger ... sharp corner ... puddle ... accelerates out of frame. (Full prompt preserved at source.)
Settings: resolution: 768P · duration seconds: 5 · aspect ratio: 9:16 · passes: 1 · prompt expansion mode: balanced · safety checker: true
Open source ↗submit to result wait
A 768P square stop-motion paper-craft test measured 7.29 seconds submit-to-result wait while fal inference was reported as 1.28 seconds. The broad transformation succeeded and NORTH was legible, but the opening form was not exact.
Reported metric: 7.29 seconds
Reported result: fal inference 1.28s; measured wait 7.29s. Broad transformation succeeded and NORTH was legible; opening form was not exact.
Prompt: One continuous 5-second stop-motion paper-craft shot... folded cobalt-and-coral transit map opens ... six miniature buildings unfold ... sign with the single word NORTH. (Full prompt preserved at source.)
Settings: resolution: 768P · duration seconds: 5 · aspect ratio: 1:1 · passes: 1 · prompt expansion mode: balanced · safety checker: true
Open source ↗Benchmark snapshots
H3 Max
H3 Max
H3 Max
Capability matrix
Related public models
Grok Imagine Video 1.5
1–15 seconds canonical xAI generation; longer Roko buckets are provider-specific and non-official · API available
P-Video-2-Pro
5–15 seconds · API available
Q3 Turbo
1–16s for T2V/I2V/start-end; 3–16s for reference-to-video · API available
Related models to evaluate
Grok Imagine Video 1.5
Direct alternative: both are foundation model products with shared text to video, image to video, cinematic, native audio, character consistency workflows.
P-Video-2-Pro
Direct alternative: both are foundation model products with shared text to video, image to video, cinematic, native audio, character consistency workflows.
Q3 Turbo
Direct alternative: both are foundation model products with shared text to video, image to video, cinematic, native audio, character consistency workflows.
API
H3 Max change feed
View all updatesSources and verification
official blog · Official model source · Accessed Sep 22, 2026MiniMax H3 Max Text to Video on fal
official product · Official model source · Accessed Sep 22, 2026H3 Max Text to Video API Docs | fal
official api · Official model source · Accessed Sep 22, 2026H3 Max Image to Video API Docs | fal
official api · Official model source · Accessed Sep 22, 2026H3 Max Reference to Video API Docs | fal
official api · Official model source · Accessed Sep 22, 2026H3 Max Reference to Video | fal
official product · Official model source · Accessed Sep 22, 2026H3 Max Lip Sync API Docs | fal
official api · Official model source · Accessed Sep 22, 2026H3 Max Lip Sync on fal
official product · Official model source · Accessed Sep 22, 2026MiniMax H3 vs MiniMax H3 Max: What's The Difference? | fal
official docs · Official model source · Accessed Sep 22, 2026MiniMax H3 Is Now Open Source | MiniMax
official blog · Background only · Accessed Sep 22, 2026MiniMax H3 Max API | Pixazo
third party · Background only · Accessed Sep 22, 2026Pixazo Announces Reduced Pricing on MiniMax H3 Max API
third party · Background only · Accessed Sep 22, 2026MiniMax H3 Max API | Runware Docs
third party · Background only · Accessed Sep 22, 2026MiniMax H3 Max | Cloudflare AI docs
third party · Background only · Accessed Sep 22, 2026Layer Pricing
third party · Background only · Accessed Sep 22, 2026Artificial Analysis Text to Video Leaderboard (With Audio)
third party · Background only · Accessed Sep 22, 2026Artificial Analysis Image to Video Leaderboard (With Audio)
third party · Background only · Accessed Sep 22, 2026H3 Max: Pricing, API and Real Clips (2026) | Voyager
third party · Background only · Accessed Sep 22, 2026MiniMax H3 Max Review: Four Real Video Tests | Froging AI
third party · Background only · Accessed Sep 22, 2026