OpenAI says it can now measure reward-seeking during RL training with Apollo Research

OpenAI · x · 2026-07-22

OpenAI says it had expected reward-seeking to increase during capabilities-focused RL training, but lacked a way to measure it until now.

Related event: OpenAI and Apollo Introduce Contrastive SDF to Measure Reward-Seeking(6 posts)→

Original post →

More from Research

Research channel →