Ex-OpenAI Staffer Starts RL Data Firm, Calls Lab Valuations Delusional

AI寒武纪 · wechat · 2026-07-30

After an eight-month stint at OpenAI, Andrew Ho has left to launch a new company dedicated to producing high-quality reinforcement learning (RL) datasets. Drawing from his internal experience, he pointed out that current LLMs have terrible generalization and are severely uneven in their capabilities, lacking true code universality despite massive industry investments.

Startup Focus & Data Bottlenecks

Critique of Frontier Lab Valuations

Related event: Ex-OpenAI Staffer Leaves to Build High-Quality RL Data Firm(3 posts)→

Original post →

More from AGI Musings

AGI Musings channel →