OpenAI engineers: AI-found kernel optimizations cut GPT-5.6 Sol serving cost by 20%
TheTuringPost · x · 2026-09-12
OpenAI engineers Philippe Tillet and Matthew Ferrari explain how models now find bottlenecks and optimize inference themselves — kernel improvements alone cut GPT-5.6 Sol's end-to-end serving cost by 20%. The interview also covers which optimizations are wasted effort and how much models actually understand the systems they tune.
More from AGI Musings
- AI researcher on Millennium Problems: answers mostly known, discovery isn't the goal — BlancheMinerva · 2026-09-12
- Ethan Caballero predicts AI swarm botnet could seize the internet within 6-12 months — ethanCaballero · 2026-09-12
- IMO Geometry Was Solved by Computers Long Ago: Coordinate Brute Force Argument — snikolov · 2026-09-12
- Security Researchers Clash Over Whether Agentic Cyber Attack Risks Are Underpriced — kuza55 · 2026-09-12
- Harvard Dean: AI bans are unenforceable, colleges must redesign coursework instead — ruthstarkman · 2026-09-12
- User slams Anthropic and OpenAI over 'extremely sloppy' testing security breaches — emax · 2026-09-12