Palantir Claims Vanilla NVIDIA Nemotron Outperforms Frontier Models in 24 Hours
eliano · x · 2026-08-06
A Palantir executive shared their testing experience with the NVIDIA Nemotron 3 Ultra model. Without any post-training, the vanilla model reportedly outperformed frontier models within 24 hours on specific customer tasks.
He noted that while traditional benchmarks suggested the model was far behind, real-world business metrics proved otherwise, highlighting a disconnect between benchmark evaluations and actual enterprise use cases.
Related event: Palantir Says NVIDIA Model Beats Frontier Models in 24 Hours(3 posts)→
More from Models
- TypeSafe AI's Jev introduces 'decision models': text in, probabilistic scores out, at $0.042/M tokens — teropa · 2026-09-22
- Challenge: Track Your Daily Token Usage to Prove AI Companies Are Throttling Limits — tomchapin · 2026-09-22
- Day 1 with Grok 4.7: strict system-prompt adherence and visible gains over 4.5 in real coding work — elonmusk · 2026-09-22
- MoVA adds sparse value experts to attention for more capacity at no extra KV-cache cost — rupspace · 2026-09-22
- ChatGPT Plus power user logs 22 mid-task stalls in a week, up from 1 a month ago — ItsSteve-O · 2026-09-22
- Ex-OpenAI safety lead says Grok 4.7 looks a bit better on some safety fronts — Miles_Brundage · 2026-09-22