Ex-Meta engineer benchmarks Jev vs GPT-nano: same score, 5x faster, 20% pricier

danielmckinn0n · x · 2026-09-25

Former Meta engineer Daniel McKinnon shared hands-on benchmarks of Jev, questioning the hype: his first LLM project in 2022 used OPT-175B as a first-line classifier, so Jev felt like "just LLM + constrained decoding."

With few public benchmarks available, he compared Jev against GPT-nano on a large-scale document classification task:

He finds the speed tradeoff worthwhile for his use case, but suspects a fine-tuned small model like Qwen or Gemma could be faster, better and cheaper for continuous workloads — and openly asks for an ELI5 on what the excitement is about.

Original post →

More from coding & agent

coding & agent channel →