Jev after days of real testing: weak Chinese, no reasoning, 30x faster than DeepSeek only in closed tasks

ZeYanjie · x · 2026-09-20

After days of heavy use, the author maps Jev's real boundaries: it only works with strictly defined constraints, not open-ended tasks.

Verdict: Jev shines in narrow, closed, latency-sensitive jobs — a mail-triage demo sorted 300 emails into 15 departments in 9.9s for $0.036, while DeepSeek read only 10 emails in the same time. The author calls for a community-built Jev harness.

Related event: Hands-on Tests Reveal Jev's Limits: Weak Chinese, Narrow Use(2 posts)→

Original post →

More from Models

Models channel →