Alibaba releases Qwen3.8-Omni-Flash, a cheap fast omni model for text, image, audio and video
gaganghotra_ · x · 2026-09-18
Alibaba has released Qwen3.8-Omni-Flash, a cheap and fast omni model that takes text, image, audio and video as input in one brain and answers in text. The company claims roughly a 25% capability improvement over the previous Omni model, and positions it as a Gemini 3.8 Flash-class swap for audio/video understanding.
Related event: Alibaba Qwen Releases Qwen3.8-Omni-Flash, an Agentic Omni-Modal Model(8 posts)→
More from Models
- An AI forecaster has won the seasonal Metaculus Cup for the first time — NathanpmYoung · 2026-09-18
- After Jev's classification model hit, will generic regression and time-series foundation models be next as a service? — hichaelmart · 2026-09-18
- Jev vs generic model at same cost: 82.9% vs 58.8% on MMLU-Pro — simonguozirui · 2026-09-18
- DeepSeek V4.1-Flash vs V4-Pro Benchmarked: 3x Cheaper but Slower and Weaker Output — Arindam_1729 · 2026-09-18
- SemiAnalysis analyzed 616,300 OpenAI responses: GPT 5.6 Terra averages 1.94 tool calls per reply — AccBalanced · 2026-09-18
- Gemini's Japanese Apologies Are So Dramatic It Sounds Like Seppuku — DigitalFossilist · 2026-09-18