New discrete-diffusion LLM speedup paper Uno called out for similarity to Orthrus
_akhaliq · x · 2026-09-09
A new Hugging Face paper page, "Unlocking Lossless Speedups in LLMs via Discrete Diffusion," presents Uno, which combines AR and diffusion generation while keeping the architecture (and its attention) unchanged for lossless speedups.
A commenter questioned whether the core framework is the same as Orthrus (arXiv:2605.12825), released four months earlier. Author Subham Sekhar Sahoo responded that Orthrus adds diffusion attention heads with bidirectional attention, changing the architecture, while Uno does not — acknowledging similarities but disputing that the core ideas match; a discussion will be added in a revision.
Related event: Uno: Discrete Diffusion Delivers Lossless LLM Speedups(4 posts)→
More from Research
- Prompt optimizer GEPA lifts Meta Muse Spark 1.1 success from 22.2% to 100% while cutting queries to 0.3% — iamrobotbear · 2026-09-09
- Why diffusion model samples drift toward the dataset center at high noise levels — YouJiacheng · 2026-09-09
- OpenAI: 10,000 coordinating AI agents solved Navier–Stokes in 88 hours — i_dg23 · 2026-09-09
- Terence Tao weighs the tradeoff: 100 solutions, 90 publication-quality writeups, 10 left behind — tak3sh8 · 2026-09-09
- Terence Tao on how new tools flatten math's difficulty landscape while expanding its frontiers — burny_tech · 2026-09-09
- UCLA professor challenges Anandkumar's Euler singularity paper: stability proof still unfinished — lpachter · 2026-09-09