Qwen3.8-Flash launches on QwenCloud with 1M context and aggressive pricing
Alibaba_Qwen · x · 2026-08-27
Qwen3.8-Flash is now available on QwenCloud, featuring a native 262K context window (extensible to 1M tokens). Pricing is set at $0.15/1M input tokens and $0.47/1M output tokens, with cache hits costing just $0.016/1M. The model supports multimodal inputs, function calling, structured outputs, web search, and fine-tuning, while maintaining compatibility with OpenAI and Anthropic API protocols.
Related event: Alibaba Releases Qwen3.8-Flash Multimodal MoE Model to Wide Acclaim(26 posts)→
More from Models
- ThursdAI Recap: GLM 5.3 and Qwen 27B Drop, OpenAI Pauses RL for Safety — altryne · 2026-08-28
- Qwen3.8-Flash-Next runs 162k context on dual 7900 XTX — BigYoSpeck · 2026-08-28
- Glitch Exposes Gemini's Internal Thoughts — Regular_Preference64 · 2026-08-28
- GLM-5.2 Introduces Monitors to Combat Reward Hacking in RL — burny_tech · 2026-08-28
- Anthropic Luna Max test shows generous limits, high speed — timpera · 2026-08-28
- User calls Grokbot 'terrible', cites missing tasks — krishnan · 2026-08-28