Qwen3.8-Omni-Flash undercuts Gemini Flash pricing while matching its multimodal benchmarks

The Decoder · rss · 2026-09-19

Alibaba's Qwen has released Qwen3.8-Omni-Flash, its first multimodal model designed for AI agents. Per The Decoder, the model processes audio and video together and can independently use tools to edit vlogs, translate clips, or summarize movies. On audio-video benchmarks it nearly matches Gemini 3.8 Flash — at a fraction of the API cost — making it a aggressively priced entry into the multimodal agent space.

Original post →

More from Models

Models channel →