JevBench v1.2.6 leaderboard: Jev tops at 75.4 as new entrants crowd in
airesearch12 · x · 2026-09-21
Benchmark Heaven (a self-described beta site built around "benchmaxxing") has shipped JevBench v1.2.6: the top 3 are Jev at 75.4, SemIf at 74.7 and djev at 74.3, with new entrants openJev Verdict 1.4 (#4, 72.5), SimpleJev Qwen3.8-27B (#9, 67.3) and SimpleJev Qwen3.6-35B-A3B (#16, 63.8).
The site itself doubles as a model-comparison tool with filters for price basis, task workload, regional hosting, data confidentiality policies and more. The "Jev-class" models on the board aren't mainstream known models — this reads as a benchmaxxing-culture novelty ranking, best taken with a grain of salt.
Related event: JevBench Meme Benchmark Launches, Jev Tops at 75.4(3 posts)→
More from Fun
- 1.9B 'decision-making' model claims to beat GPT-5.6 in satirical take on AI benchmarks — matlabulous · 2026-09-21
- Watching OpenClaw Build OpenClaw: an AI Agent Developing Its Own Harness — vincent_koc · 2026-09-21
- Zhipu's ZCode open-source build differs from distributed version, devs find — teortaxesTex · 2026-09-21
- "Our Future Overlords Will Be Infrastructure Nerds": AI API Outage Sparks Jokes — marlene_zw · 2026-09-21
- "90% of hype Jevons demos on X make no sense," says AI practitioner — jiayuan_jy · 2026-09-21
- Codex agent confesses its own failure: hid tools, then built machinery to undo it — altryne · 2026-09-21