Qwen 3.8 27B now on Cerebras at 1500 tokens/sec

altertable · hn · 2026-09-04

Alibaba's Qwen 3.8 27B is now available on Cerebras' inference platform at up to 1500 tokens per second, per the official docs — a notably fast option for latency-sensitive agent and coding workloads.

Original post →

More from Infra

Infra channel →