Qwen3.8-Max launches on DigitalOcean Serverless with 1M context
Alibaba_Qwen · x · 2026-08-14
Alibaba Cloud's Qwen3.8-Max (2.4T params, 95B active) is now available on DigitalOcean Serverless Inference, powered by NVIDIA HGX B300 GPUs, with 1M context for long-horizon coding. Usage-based pricing, no infra management.
More from Models
- Qwen3.8-27B-FP8 on GH200: 10 concurrent streaming requests, first token in 10ms — MaziyarPanahi · 2026-08-15
- Alibaba Releases Qwen3.8-27B: 27B Parameters, Opus 4.6-Class Performance, Runs Locally — haider1 · 2026-08-15
- Qwen3.8-27B Open Source: 206 tok/s on RTX 5090 with SGLang Day-0 Support — cedric_chee · 2026-08-15
- Qwen Official: Qwen3.8-27B Significant Jump, Try It Out — Alibaba_Qwen · 2026-08-15
- Opus 4.6 Max Now Runs Locally, Enabling Cutting-Edge Research Anywhere — rand_longevity · 2026-08-15
- Hugging Face report: small models dominate real-world usage, Qwen leads local inference — huggingface · 2026-08-15