Judgment-only model Jev opens to all with $5 free credit, sparking calibration debates

机器之心 · wechat · 2026-09-21

TypeSafeAI has removed Jev's waitlist, opening the structured-judgment "SystemOne model" to everyone with $5 (120M tokens) of free credit. Developers feed it a state plus a question with fixed options, and it returns a choice, a probability, or an ordered score that code can consume directly — suited for model routing, moderation triage, and agent tool-call risk checks.

Praised generalization: Sebastian Raschka notes Jev accepts natural-language labels at runtime, avoiding per-label-set retraining that plagues traditional encoders; he suspects the secret is data and API design rather than the training algorithm. An unconfirmed report says it was post-trained on fully synthetic data atop an existing model.

Reliability questions: users report inconsistent probabilities across identical prompts and heavy order sensitivity; @predictaddict found Jev's calibration trails CatBoost across eight real datasets. TypeSafeAI itself flags instability on arithmetic, counting, and adversarial inputs. One user's Jev-powered trading bot has lost $31,680.

Related event: Ex-OpenAI researcher launches Jev, a System 1 decision model, now open to all(14 posts)→

Original post →

More from coding & agent

coding & agent channel →