SOOFI appears to have seen several of the benchmarks used against Nemotron
JJitsev · x · 2026-07-26
The thread links out to the datasets used in training and compares them with the benchmarks used for evaluation.
- The author says many evals in the SOOFI dataset came from TU Darmstadt.
- He contrasts them with the datasets used to train Nemotron 3 Nano.
- The key point is that several benchmarks used in the comparison were already seen by SOOFI but not by Nemotron.
More from Models
- Opus-5 is getting attention for its unusual vocabulary choices — adonis_singh · 2026-07-26
- Google’s Gemma team asks what capabilities people want in the next models — osanseviero · 2026-07-26
- Alibaba answers with Qwen3.8 as Kimi K3 and Chinese models keep closing the gap — emmanuelvivier · 2026-07-26
- Moonshot AI launches open-weight Kimi K3 and claims strong results against top US models — emmanuelvivier · 2026-07-26
- Silicon Valley is split on Chinese open-weight models now rivaling top U.S. systems — emmanuelvivier · 2026-07-26
- Kimi-K3 claims a perfect 6/6 on IMO 2026 Lean 4 proofs — songhan_mit · 2026-07-26