Alex Zhang: design the language model's shape around agents, with recurrent memory for old history

CatAstro_Piyush · x · 2026-10-01

Stanford researcher Alex L. Zhang published a blog post, "Language Model 'Shape'", arguing the input/output shape of LLMs has been static since ChatGPT, and all agent work fits a harness around the autoregressive decoder-only Transformer. He asks whether it's worth flipping that: change the model's shape to fit the harness.

Key points:

Original post →

More from coding & agent

coding & agent channel →