Datology AI: Curated Pretraining Data Lifts 30B-A3B to 46.8% vs 37.7% Baseline

josh_wills · x · 2026-10-09

Datology AI reports that its curated pretraining data pushed a 30B-A3B model to 46.8% average benchmark score vs. 37.7% for the baseline, with gains persisting after post-training. A curated 12B-A1.4B model beat the 30B-A3B baseline despite using roughly one-fifth of the pretraining compute.

Original post →

More from Research

Research channel →