Reddit speculation: labs may train 20T-param internal-only models never served publicly
Crazyscientist1024 · reddit · 2026-09-28
A Reddit post speculates that the latest model generation (referred to as Astra and Fable) is already dramatically speeding up labs' internal R&D, making it rational to train 20T-parameter models that are too expensive to serve publicly (likely hundreds of dollars per million tokens) but valuable purely as internal R&D accelerators — a self-reinforcing loop where AI builds better AI, producing giant models the public never sees.
More from AGI Musings
- RL's limit: without a scoring function there's no signal, and 'learning' is inflated jargon — gerardsans · 2026-09-28
- Mark Pincus: If starting over, he'd rebuild social media with AI agents — PeterDiamandis · 2026-09-28
- VC admits he was wrong on 'token apocalypse' as Opus 5.5 points to too-cheap-to-meter AI — StewartalsopIII · 2026-09-28
- Matt Turck: AI researchers don't buy doom or acceleration, outsiders do — mattturck · 2026-09-28
- Philosopher Jeff Sebo Pushes Back on Pinker's Sorites Argument Against Superintelligence — burny_tech · 2026-09-28
- Prediction: In ~6 months AI animations will beat all but the best hand-made work — round · 2026-09-28