Jev, built by an RLHF/InstructGPT veteran, targets machine decisions with new RLCD training

MaryamMiradi · x · 2026-09-21

TypeSafe, founded by Diogo Almeida (who worked on the RLHF/InstructGPT lineage behind ChatGPT), released Jev — a model built for software to make decisions with, not for humans to chat with. Its RLCD (Reinforcement Learning for Calibrated Decisions) method optimizes for calibrated machine-consumable decisions instead of human-preferred responses. A 6-step explainer thread.

Original post →

More from Models

Models channel →