Someone built JevBench, the first benchmark for 'Jev-class models'; original Jev leads at 75.3
airesearch12 · x · 2026-09-19
Riding the hype around JEV, the tiny local computer-use model, someone launched JevBench — billed as the first benchmark for 'Jev-class models.' The original Jev by @typesafeai leads at 75.3, with SemIf close behind at 74.6. Largely a community in-joke, it playfully catalogues the wave of small, fast, screenshot-free local agent models.
More from Fun
- AI Videos Reveal a Weird Untapped Market: Backyard Underground Housing — generativist · 2026-09-19
- OpenAI's credit rating nudged toward investment grade as labs warn of existential AI risk — AlexTensor · 2026-09-19
- Researcher uses AI to illustrate how AI writing is breaking peer review — sethlazar · 2026-09-19
- Weekend project plans ruined: developer vents about favorite model getting nerfed — BrandNewFeel · 2026-09-19
- One Paper Per Lifetime? Bronstein Claps Back With a Parrot Joke — mmbronstein · 2026-09-19
- :-) Emoticon Turns 44; Inventor Fahlman Asks CMU Faculty "Do You Work on LLMs?" — aran_nayebi · 2026-09-19