Together AI Launches Serverless Inference for Qwen3.8-2.4T-A95B
Together AI has launched serverless inference support for the Qwen3.8-2.4T-A95B model, featuring 256K context and optimized for high-throughput coding and agentic workloads, enabling robust long-range task management.
2026-08-13 ~ 2026-08-13 · 2 related posts
- Episode 1: Alibaba Open-Sources 2.4T Parameter Flagship Qwen3.8-Max(2026-08-12, 17 posts)
- Episode 2: vLLM Announces Day-0 Support for Qwen3.8 2.4T Model(2026-08-13, 2 posts)
- Episode 3: Together AI Launches Serverless Inference for Qwen3.8-2.4T-A95B(2026-08-13, 2 posts)
- Qwen3.8-2.4T-A95B for long-running agents: 256K context, self-testing, function calling — togethercompute · 2026-08-13
- Together AI launches serverless inference for Qwen3.8-2.4T-A95B with 99.9% SLA — togethercompute · 2026-08-13