Marin 535B open-source model training starts with 2.7e24 FLOPs plan

DavidBennett__ · x · 2026-08-24

Percy Liang announced the start of training for the Marin 535B-A23B model. The project is fully open, including code, data, recipes, and experimental results. Training will run on 11 GB200 NVL72 clusters for approximately 3 months, involving a total compute of 2.7e24 FLOPs across 18.75T tokens for pre-training and mid-training. The team previously debugged and forecasted the run using a scaling ladder from 1.6B to 27.7B models.

Original post →

More from Infra

Infra channel →