FLARE: Open-Sourced Diffusion LLM Boasts Near GPT-5 Performance
bodonoghue85 · x · 2026-08-14
Adobe Research and Georgia Tech have introduced FLARE, a diffusion language model. It claims to achieve near GPT-5 performance while decoding significantly faster—up to 4.8x faster than other open diffusion LLMs.
FLARE presents a systematic conversion framework that transforms strong autoregressive (AR) hybrid-attention LLMs into diffusion LLMs. With a single, low-cost training stage, one checkpoint supports two decoding modes: AR-style speculative decoding for quality, and diffusion-style parallel denoising for speed. Models at 2B, 4B, and 9B parameters have been fully open-sourced.
More from Models
- Vercel Offers GLM 5.2 Model Free for eve Agents Until August 27 — cramforce · 2026-08-14
- Deepgram Crosses $100M ARR and Launches Flux TTS Voice Model — deepgramscott · 2026-08-14
- Musk Offers More Free Usage and Resets Limits for Grok 4.6 Launch — EricBuess · 2026-08-14
- a16z's Martin Casado Tests Grok 4.6: Impressed by Complex Coding and Long Tasks — elonmusk · 2026-08-14
- Frontier LLM Token Prices: A Reflection of Underlying Model Sizes — sergeykarayev · 2026-08-14
- Meta Releases Muse Glimmer: A 30B Local Agent Model — ollama · 2026-08-14