Gemini Omni Flash Tops Video Leaderboards
ArtificialAnlys · x · 2026-07-14
Google's Gemini Omni Flash has taken the top spot on Artificial Analysis's Text-to-Video and Image-to-Video leaderboards, slightly edging out ByteDance's Seedance 2.0.
Key model highlights:
- Natively multimodal, supporting text/image/video inputs.
- Can generate videos with native audio and supports conversational editing.
- Output specs: 3–10 seconds, 720p, 24 FPS, supporting 16:9 / 9:16.
- Pricing: $0.10/second (or $6/minute), on par with Veo 3.1 Fast.
- Now available via Gemini API, Google AI Studio, Gemini Enterprise Agent Platform, Gemini App, and Google Flow; free to use in YouTube Shorts and YouTube Create.
Related event: Google Gemini Omni Flash Tops Video Generation Leaderboards(2 posts)→
More from Multimodal
- Dev builds interactive 3D product experience with GPT-6 Astra + Hyper3D Rodin — nikola_mr64990 · 2026-09-11
- Using a finisher move on one mosquito with MiniMax H3 MAX — the bug survives — Hailuo_AI · 2026-09-11
- Skyfall GS Uses Flux to Refine Gaussian Splatting, Accepted at ECCV 2026 — ducha_aiki · 2026-09-11
- Lumara AI Film Festival Comes to NYC Oct 26, Top AI Filmmakers to Compete — 0xAllen_ · 2026-09-11
- Pterodactyl Detective: An AI-Generated Proof-of-Concept Trailer — PterodactylDetective · 2026-09-11
- Imperium Game Trailer Showcases AI Video Generation — keaslenyt · 2026-09-11