Gemini Omni Flash Tops Video Leaderboards
ArtificialAnlys · x · 2026-07-14
Google's Gemini Omni Flash has taken the top spot on Artificial Analysis's Text-to-Video and Image-to-Video leaderboards, slightly edging out ByteDance's Seedance 2.0.
Key model highlights:
- Natively multimodal, supporting text/image/video inputs.
- Can generate videos with native audio and supports conversational editing.
- Output specs: 3–10 seconds, 720p, 24 FPS, supporting 16:9 / 9:16.
- Pricing: $0.10/second (or $6/minute), on par with Veo 3.1 Fast.
- Now available via Gemini API, Google AI Studio, Gemini Enterprise Agent Platform, Gemini App, and Google Flow; free to use in YouTube Shorts and YouTube Create.
Related event: Google Gemini Omni Flash Tops Video Generation Leaderboards(2 posts)→
More from Multimodal
- Seedance 2.0 demo turns ketchup on spaghetti in Rome into an AI reaction meme — azed_ai · 2026-07-21
- A reusable “Lunar Eclipse Dreamscape” prompt comes with multiple example renders — LudovicCreator · 2026-07-21
- Midjourney 8.2 preview shows a double-exposure prompt with strong style control — michaelrabone · 2026-07-21
- Travel MCP Server adds flight, hotel, weather and budget tools for agents — modelcontextprotocol · 2026-07-21
- Douyin Video Analysis MCP turns share links into structured video summaries — modelcontextprotocol · 2026-07-21
- Synthesia launches Dubbing 2.0 with 130+ languages and lip-sync video translation — synthesiaIO · 2026-07-21