Mistral Trains 1T-Parameter Model on 3,800 Grace Blackwell Chips in Europe
Mistral trained its 1-trillion-parameter Mistral Large 4 end-to-end (pretraining to post-training) on 3,800 NVIDIA Grace Blackwell superchips at its European data center, funded by a €3B Series D, with two more clusters on the way.
2026-10-06 ~ 2026-10-06 · 3 related posts
- ML4 Trained on 3,800 Grace Blackwell GPUs; Mistral's C and D Round Clusters Coming Online — GuillaumeLample · 2026-10-06
- 1T-parameter model fully pre- and post-trained on 3,800 Grace Blackwells in Europe — qtnx_ · 2026-10-06
- Mistral Large 4 trained on nearly 4,000 NVIDIA Grace Blackwell Superchips, backed by €3B Series D — cedric_chee · 2026-10-06