Study Finds Faithful World Models Inside Transformers

An arXiv paper on taxiGPT, trained on Manhattan random-walk data, found a faithful internal world map, challenging the claim that Transformers lack world models. Feature superposition interferes with, but does not eliminate, the learned world model.

2026-09-23 ~ 2026-09-23 · 2 related posts