DeepMind Releases DiffusionGemma: Discrete Diffusion for Ultra-Fast Text Generation

deepmind · hf · 2026-08-04

Google DeepMind released the technical report for DiffusionGemma, an experimental open-weight language model. Unlike conventional autoregressive (AR) models that decode token-by-token, it uses discrete diffusion to iteratively refine blocks of 256 tokens in parallel, avoiding the sequential decoding bottleneck.

Key Highlights:

Original post →

More from Models

Models channel →