Rob Leclerc: model unmonitorability is a lab choice, not an inevitability

robleclerc · x · 2026-09-18

Responding to complaints that models are becoming unmonitorable, Rob Leclerc argues that training for form over function effectively relaxes constraints on legibility — so the trend should be expected. But he pushes back on framing it as inevitable: adding legibility constraints would increase legibility at some cost to model performance. He calls it dangerous to imply that as models get more powerful, opacity is simply the way things will be. It's a design choice wholly under the control of the labs, not a law of nature.

Original post →

More from AGI Musings

AGI Musings channel →