Hands-on: TypeSafe's tiny Jev classifier goes head-to-head with Claude Haiku/Sonnet/Opus on text understanding

vesko_st · x · 2026-09-22

Ves Stoyanov (Lightfield) tested TypeSafe's newly released Jev, a compact "System One Choice" classifier that takes a passage, a question, and options and returns one answer. Wary that public benchmarks may be contaminated and stale, he re-ran everything under a fixed protocol and hand-wrote a corpus as a contamination hedge. Jev was benchmarked against Claude Haiku 4.5, Sonnet 5, and Opus 5 on CommonsenseQA, MMLU-CF, RACE-H, Grace and more, priced per 1,000 long-context questions. His takeaway: the purpose-built small classifier now holds its own on text understanding, making dedicated classifier models compelling again for routing-style tasks.

Related event: Tiny classifier Jev matches Claude Sonnet at ~150x lower cost(2 posts)→

Original post →

More from Models

Models channel →