JevBench v1.3.0 Released: Original Jev Still Leads at 74.4, 47 Rivals Closing In
airesearch12 · x · 2026-09-22
JevBench has been updated to v1.3.0. The original Jev remains #1 with a score of 74.4, but the leaderboard now includes 47 competing models, some of which are getting very close. The benchmark is hosted on Benchmark Heaven, which supports filtering and comparing models by price basis (input/output blends), regional hosting, and data confidentiality policies.
More from Models
- OpenAI criticized for claiming 100 open math problems solved without disclosing the total attempted — burny_tech · 2026-09-22
- Dev argues prompting alone can't get AI to solve natural science problems — felpix_ · 2026-09-22
- Dev claims benchmarks are 'absolutely meaningless' — models only differ by vibe — gnukeith · 2026-09-22
- Report: OpenAI Trained New Math Model in ~11 Days, Solved Longstanding Open Problems — felpix_ · 2026-09-22
- Jev Hallucinates at Rates Similar to Other Models, Developer's Examples Show — JeremyNguyenPhD · 2026-09-22
- Aikido Security releases Altar-1, its first open-weight security model — HankYeomans · 2026-09-22