Celeris-1 Launches with Diffusion-Based Low-Latency Inference

Celeris Labs has launched the Celeris-1 language model, utilizing a diffusion-based inference architecture to achieve a 157ms p50 latency. The model claims to offer near GPT-5 level intelligence while significantly improving response speed.

2026-07-24 ~ 2026-07-25 · 3 related posts