AI circles clash over CoT monitorability: HF attack probe hinged on chain-of-thought

tomekkorbak · x · 2026-09-03

A debate erupted over rumors that OpenAI is pursuing "neuralese" and abandoning chain-of-thought monitorability:

Related event: OpenAI researchers push back on neuralese fears, saying frontier models remain monitorable(20 posts)→

Original post →

More from Safety

Safety channel →