SOOFI’s own table shows weaker scores once you remove training-set benchmarks

JJitsev · x · 2026-07-26

The author uses the report’s own Table 5 to show how training on eval sets shifts the comparison.

Related event: SOOFI Model Accused of Severe Benchmark Contamination in Comparison with Nemotron(9 posts)→

Original post →

More from Models

Models channel →