Qwen3.8-Flash launches on QwenCloud with 1M context and aggressive pricing

Alibaba_Qwen · x · 2026-08-27

Qwen3.8-Flash is now available on QwenCloud, featuring a native 262K context window (extensible to 1M tokens). Pricing is set at $0.15/1M input tokens and $0.47/1M output tokens, with cache hits costing just $0.016/1M. The model supports multimodal inputs, function calling, structured outputs, web search, and fine-tuning, while maintaining compatibility with OpenAI and Anthropic API protocols.

Related event: Alibaba Releases Qwen3.8-Flash Multimodal MoE Model to Wide Acclaim(26 posts)→

Original post →

More from Models

Models channel →