Gacha Decoding diversifies LLM outputs with 11x better sample efficiency

natolambert · x · 2026-10-03

Researchers introduce Gacha Decoding, a new method for diversifying LLM responses that beats prior work by over 11x in sample efficiency. Instead of relying on token entropy, it leverages the model's instruction-following ability plus an external RNG — reroll the LM like a gacha game to get varied ideas, synthetic data, or RL environments.

Original post →

More from Research

Research channel →