ML4 Trained on 3,800 Grace Blackwell GPUs; Mistral's C and D Round Clusters Coming Online
GuillaumeLample · x · 2026-10-06
Part 2 of the ML4 thread: the model was trained on 3,800 NVIDIA Grace Blackwell GPUs at Mistral's Bruyères-le-Châtel datacenter in Europe, built with Series B funds. Series C and D clusters are coming online, enabling longer training, more ambitious post-training, and faster iteration, with large and rapid improvements expected in the weeks and months ahead.
Related event: 1T-Parameter Model Trained Fully in Europe on 3,800 Grace Blackwell GPUs(2 posts)→
More from Infra
- Google Buys 890 MW of Nuclear Without Building a Single New Reactor — MicahBerkley · 2026-10-06
- Cycle.io Launches DevOps MCP: 3-Node Mongo Replica Set Across 3 Clouds in 15 Minutes — AlexMattoni · 2026-10-06
- Weaviate ships query profiling: a 48ms slow query turned out to be disk reads, not HNSW — CShorten30 · 2026-10-06
- HN Debate: Did Oracle Just Trigger the Implosion of the AI Bubble? — mpweiher · 2026-10-06
- Lambda adopts NVIDIA's AIPerf for model cards showing real-workload inference benchmarks — TheZachMueller · 2026-10-06
- Nvidia nears $6 trillion market value as AI frenzy keeps pushing stocks higher — AryHHAry · 2026-10-06