OpenAI launches inaugural Safety Fellows cohort on alignment and interpretability
tomekkorbak · x · 2026-09-17
OpenAI has kicked off its inaugural Safety Fellows cohort in partnership with Constellation. First-cohort member @xaverg announced the group will spend the next few months working on core issues in AI alignment, control, and interpretability, with support from mentors and the broader community.
DeepMind researcher Tomáš Korbak shared the news, saying he expects "an unreasonable amount of ground-breaking AI safety research" to come out of this group.
More from Companies & People
- Databricks rolls out Astra to 3,500 engineers, sees 60% coding spend increase — pwendell · 2026-09-17
- Musk spotlights live experiment: three employees building a company in 3 days with Grok — elonmusk · 2026-09-17
- Spotify's AI Persona Badge Targets Photorealistic Identities, Not Royalties — mixtapedmonk · 2026-09-17
- OpenAI Devs Team With Product Hunt on GPT-6 Astra Challenge: $10K API Credits — OpenAIDevs · 2026-09-17
- Microsoft's Suleyman: Anthropic's model-welfare training could make Claude harder to control — rohanpaul_ai · 2026-09-17
- Jensen Huang at All-In Summit: AI leadership will be built by everyone, open and closed models both matter — NVIDIAAI · 2026-09-17