TabPFN-3.5 tops Kaggle's Otto competition out of the box, but experts call the benchmark flawed

RichmanRonald · x · 2026-09-16

TabPFN-3.5 launched today claiming rank 1 on the historic Kaggle Otto competition with no tuning — innixma argues this is more impressive than anything in the report, showing hidden signal in the data.

JFPuget pushed back: results on past Kaggle competitions mean little, for at least 3 reasons:

A representative debate on whether historical competition results make a valid benchmark for new models.

Original post →

More from Models

Models channel →