Balesni warns recurrent LLMs would deal a huge blow to safety

j_asminewang · x · 2026-09-02

Balesni argues that switching to fully recurrent LLM architectures would be a major blow to AI safety history. He urges labs to limit the opaque serial depth of models to avoid worst-case scenarios. OpenAI researcher Merettm countered that current frontier models' computation graph depth is within a factor of two of GPT-4, emphasizing OpenAI's commitment to chain-of-thought monitoring for alignment generalization.

Related event: OpenAI Researchers Push Back on Neuralese Fears: Frontier Models Remain Monitorable(11 posts)→

Original post →

More from Safety

Safety channel →