OpenAI Details New Training Methods for Non-Verifiable Domains in Astra

morqon · x · 2026-09-05

The quoted post from OpenAI's kevinwng says Astra made significant progress on subjective domains like design aesthetics and knowledge work, enabled by new methods for training on non-verifiable domains — tasks without objective ground truth where standard RL is hard to apply directly. He calls this line of work exciting, teases more to come, and invites users to try Astra.

Original post →

More from Models

Models channel →