Ex-OpenAI Staffer Leaves to Build High-Quality RL Data Firm
Former OpenAI employee Andrew Ho has left to start a company focused on high-quality reinforcement learning (RL) datasets. He noted that despite billions in investments, current LLMs suffer from poor generalization, arguing that high-quality RL data is the key to overcoming industry bottlenecks.
2026-07-30 ~ 2026-07-30 · 3 related posts
- Ex-OpenAI Staffer Leaves to Build High-Quality RL Data Startup, Cites Poor LLM Generalization — panickssery · 2026-07-30
- Ex-OpenAI Staffer Starts RL Data Firm, Calls Lab Valuations Delusional — AI寒武纪 · 2026-07-30
1 near-duplicate retellings: rickasaurus