Open-weight models near 3T parameters as inference emerges as the next moat

ruslansv · x · 2026-07-21

The post argues that open-weight models are already approaching 3T parameters, while a rumored Fable model may have 6–8T parameters. If that trajectory continues, the author expects 20T-scale models by year-end.

The key claim is that inference on such colossal models could become the main moat. The post also notes that training corpora are measured in tens of trillions of tokens, implying that compute-efficient inference may matter more as models and datasets keep scaling.

Related event: Open-Source Models Approach 3T Parameters, Inference Compute Becomes the New Moat(3 posts)→

Original post →

More from Research

Research channel →