DRIFT: A New Framework from Tsinghua and Beike for Continuous LLM Self-Evolution

青稞AI · wechat · 2026-07-26

A joint research team from Tsinghua University, Beike, and ENS Paris-Saclay introduced DRIFT, an online self-evolution post-training framework for LLMs. It aims to solve the challenge of how models can continuously improve after RL without collapsing.

The framework features several core mechanisms:

Experiments show that DRIFT enables self-correction and continuous evolution without external expert supervision, achieving new SOTA on multiple complex reasoning and Tool Use benchmarks. The authors will host an upcoming webinar to dive deep into the algorithm's design and theoretical motivations.

Original post →

More from Research

Research channel →