Qwen3.8-2.4T-A95B Available on Modal with 1M Token Context
AAAzzam · x · 2026-08-13
Alibaba's Qwen3.8-2.4T-A95B model is now available on the Modal platform. It is served with a custom DFlash speculator trained on tool-call-heavy data and supports the full 1 million token context window.
More from Models
- Pure Rust Browser Port: Qwen3-TTS Runs Locally Without GPU — doodlestein · 2026-08-13
- AT&T Consumes 45B Tokens Daily, Cuts Costs by 90% Shifting to Open Models — SuB8u · 2026-08-13
- Grok 4.6 and Qwen3.8-Max Released: Podcast Explores the Open vs. Closed AI Frontier — thursdai_pod · 2026-08-13
- DeepSeek V4-0813 Shows No Gains Over July Version in Benchmarks — teortaxesTex · 2026-08-13
- Conspiracy Theory on OpenAI's Next Model Astra: Intelligent Cloud System — haider1 · 2026-08-13
- Grok-4.6 Tested: Agentic Coding Benchmark Score Improves by ~5% — karminski3 · 2026-08-13