Researcher's rant on brittle frontier models sparks resonance

On September 20, a researcher posting as lateinteraction published a long tweet as a "checkpoint" on the current moment: at the "Ultra" tier, he squeezes "barely acceptable" work output from frontier models every day, and it's painful—he has to provide more detailed context and finer-grained feedback than he has ever given any human colleague in his life, and his day job is literally writing feedback for models. He described 2026's frontier models as having a "jagged capabilities" problem: even after 8 straight hours using Astra/Fable-style agentic products, as long as the work demands quality across high dimensions (more than one or two easily verifiable dimensions), the models remain as fragile as autocomplete, with no improvement across generations.

Confirmed

Why it matters

Note: Multiple posts give conflicting descriptions of lateinteraction's real identity (one says Voyage AI founder Tengyu Ma, others say ColBERT author Omar Khattab or Charles Packer); the sources contradict each other and the identity cannot be confirmed.

2026-09-20 ~ 2026-09-20 · 5 related posts

Primary sources