Westlake AGI Lab's Code World Model splits world rules (code) from video rendering

jiqizhixin · x · 2026-09-13

Westlake University's AGI Lab introduces Code World Model, addressing a key flaw in video-based training: video records the visible result of world evolution but not the rules and mechanisms producing it. Models must reverse-engineer collision, attack ranges, faction relations, and quest state from sparse visual consequences, while long-timescale changes happen off-camera—scaling video training adds seen results, not the white-box mechanics.

The method splits "how the world evolves" from "how the world is seen": a coding agent decides world evolution, code keeps rules executing persistently, and a video model converts the evolved state into high-fidelity visual observations.

Original post →

More from Models

Models channel →