Liquid AI open-sources 'antidoom' FTPO training to fix small-model doom loops

helloiamleonie · x · 2026-09-15

Leonie of Liquid AI explains that small models with thinking capabilities get stuck in "doom loops" on complex tasks. Liquid AI's fix: "antidoom" training via Final Token Preference Optimization (FTPO).

They open-sourced a full tutorial notebook showing what doom loops are, how FTPO reduces them, and how to implement a custom DPOTrainer in TRL — available in the Liquid4All/cookbook repo (2.5k stars).

Related event: Liquid AI Open-Sources FTPO to Fix Small Models' Doom Loops(2 posts)→

Original post →

More from coding & agent

coding & agent channel →