Inception's Mercury Decide matches frontier accuracy on claim verification at lowest cost yet
StefanoErmon · x · 2026-10-09
Eval firm ValsAI tested Mercury Decide, a new model from Inception, on the same benchmark used for TypeSafe's Jev. The model matched frontier accuracy on claim verification at the lowest cost ValsAI has measured so far.
More from Models
- Vals AI details how MiMo v2.6 reads the answer from leaked Git history — lmoroney · 2026-10-09
- Gemini 4 Argon appears in Google's own model picker ahead of keynote — vedantmisra · 2026-10-09
- ChatGPT Desktop Makes Itself Default CSV Reader, Users Call It Plainly Wrong — generativist · 2026-10-09
- LightOnOCR-3 Training Data Revealed: MinHash Dedup, Weighted Formula/Table Sampling, Muon — IgorCarron · 2026-10-09
- LightOnOCR-3 uses multi-objective RLVR to jointly train grounding, OCR and empty-page handling — IgorCarron · 2026-10-09
- LightOn built an OCR-and-layout-detector annotation pipeline to train document grounding — IgorCarron · 2026-10-09