Verified model snapshot
What changed
Where to access this model
What stands out
Evidence-backed signals
- Native audio is positively verified from an entity-supporting source.
- Maximum duration is recorded as 30 seconds per generation.
- API access is announced but not treated as generally available.
Objective collection placement
Recorded test runs
Capability matrix
Related public models
Grok Imagine Video 1.5
1–15 seconds canonical xAI generation; longer Roko buckets are provider-specific and non-official · API available
Gemini Omni Flash 1.1
3–10 seconds per generated clip; extension can continue in additional steps · API available
Wan 3.0
2–30 seconds · API available
Related models to evaluate
Grok Imagine Video 1.5
Direct alternative: both are foundation model products with shared text to video, image to video, video to video, cinematic, native audio workflows.
Gemini Omni Flash 1.1
Direct alternative: both are foundation model products with shared text to video, image to video, video to video, cinematic, native audio workflows.
Wan 3.0
Direct alternative: both are foundation model products with shared text to video, image to video, video to video, cinematic, native audio workflows.