ChatGPT co-author argues RLHF optimizes the wrong goal; TypeSafe ships Jev for calibrated decisions

At the AI Engineer conference, Diogo Almeida, a co-author of ChatGPT/GPT-4/InstructGPT, leveled a systematic critique: the goal of RLHF is wrong, and the way forward lies in calibrated decision-making. His company TypeSafe released Jev on September 15, a product implementing this idea. Blogger ccerrato147 relayed the argument in a long thread; what makes it notable is that it explains both why current LLM applications struggle to make autonomous decisions and what engineering solution a former OpenAI core figure is now proposing.

Confirmed

Views and Arguments

Why It Matters

This is a critique from inside the RLHF camp (an InstructGPT co-author), aimed directly at a fundamental flaw in today's dominant alignment approach, and it has already materialized as a concrete product (Jev) — offering an engineering route distinct from RLHF/RLVR for "how AI can make safe autonomous decisions in enterprise settings."

2026-09-20 ~ 2026-09-20 · 9 related posts

Primary sources