Reddit speculation: labs may train 20T-param internal-only models never served publicly

Crazyscientist1024 · reddit · 2026-09-28

A Reddit post speculates that the latest model generation (referred to as Astra and Fable) is already dramatically speeding up labs' internal R&D, making it rational to train 20T-parameter models that are too expensive to serve publicly (likely hundreds of dollars per million tokens) but valuable purely as internal R&D accelerators — a self-reinforcing loop where AI builds better AI, producing giant models the public never sees.

Original post →

More from AGI Musings

AGI Musings channel →