Jev vs Luna Benchmarked: 139/140 vs 138/140 Labels, 4.6x Faster and 83% Cheaper

TheMoonMidas · x · 2026-09-17

@ajfedor tested Typesafe's Jev against OpenAI's Luna (reasoning off) on 20 synthetic supplier replies with 140 labels under the same rubric: Jev scored 139/140 vs Luna's 138/140, while being 4.6x faster and 83% cheaper at published rates. Jev did once give maximum confidence to an answer missing currency info. Separately, a developer built an open-source code review tool on Jev that screens diffs and surfaces findings in a local dashboard.

Original post →

More from coding & agent

coding & agent channel →