OpenAI alignment evals lead: safety work driven by teams without "safety" in their names

yanndubs · x · 2026-09-13

shared a quote from OpenAI's Sam Arnesen, who leads alignment evals work: many training efforts improving safety were enthusiastically driven by teams without "safety" or "alignment" in their names. Posttraining (Yann Dubois' team) and earlier pipeline teams work closely with alignment/safety, showing safety is embedded across model development rather than siloed.

Original post →

More from Companies & People

Companies & People channel →