TRL v1.15 ships fused LM head: 82% less peak VRAM, 7x longer sequences

QGallouedec · x · 2026-10-09

Hugging Face's TRL v1.15 is out, billed as its biggest optimization ever. The fused LM head, on by default, cuts peak VRAM by up to 82% and supports 7x longer sequences: DPO goes from 10k to 59k tokens, GRPO from 29k to 115k.

Original post →

More from Infra

Infra channel →