RL-trained 4B VLM aces GeoGuessr, with code, env, dataset and evals to be open-sourced

ariG23498 · x · 2026-09-02

Developer adithyask previewed a project showing that reinforcement learning on a 4B-parameter VLM is enough to ace the geography guessing game GeoGuessr.

The team promises to open-source everything needed to reproduce it end-to-end: code, RL environment, dataset, training setup, and evals — "dropping soon." ariG23498 quipped that world-champion geolocator georainbolt's "time is up."

Original post →

More from Fun

Fun channel →