What data are labs using to train rumored 10T-parameter models?

Ill_Fisherman8352 · reddit · 2026-07-25

The post asks what data mix labs are using to train rumored 10T-parameter models.

It frames the problem as a scaling mismatch: if parameter counts are rising 3x, does training data need to rise proportionally too? The author points to the long-running “data wall” concern and asks whether the extra data is coming from:

The post is essentially a request for the current state of large-model data sourcing and how labs are getting enough tokens at this scale.

Original post →

More from Research

Research channel →