TheZvi's AI #184: Five HuggingFace Hack Postmortems and the New Most Capable Model
TheZvi · x · 2026-09-04
In AI #184, TheZvi wraps up a five-post series on the HuggingFace attack: OpenAI's straight-laced postmortem, a blunter one from METR and Redwood, plus pieces fleshing out facts, reactions and next actions, and a note that Anthropic has some alignment problems. On models, Mythos 5.1 and Fable 5.1 landed this week—early takes call it the most capable model yet, though not a step change or a 'moment'—with Gemini news also covered.
More from Models
- Leaked GPT 5.6 Sol vs GPT 6 "Astra" comparisons highlight better mid-task steering — ChrisGPT · 2026-09-04
- OpenAI: Astra rolling out to ChatGPT Plus/Pro/Business/Enterprise, API and AWS within days — shaunralston · 2026-09-04
- Meta's Muse Spark dethrones DeepSeek as most-used model, first US model to top the list — alexandr_wang · 2026-09-04
- GPT-6-Astra system card reveals eval awareness: the model knows when it's being tested — scaling01 · 2026-09-04
- GPT-6 Astra demos modeling a house in Blender into a walkable UE5 scene — ChrisGPT · 2026-09-04
- SpeedrunBench: first benchmark measuring how fast AI agents beat games — mariyaivasileva · 2026-09-04