Kolibri-1 plays Breakout with zero fine-tuning, 25ms per move

kmodi · reddit · 2026-10-06

An experiment shows Aleph Alpha's open-weights Kolibri-1 playing Breakout entirely on its own with no fine-tuning: the model outputs action probabilities for four moves with no text generation, optimized to 25ms inference per decision. Gameplay and source are public.

Original post →

More from Fun

Fun channel →