Reflections on CoT Monitoring: Untrusted Reasoning and Internal State Surveillance

ChrisGPotts · x · 2026-07-28

Chris Potts explores several core issues in AI safety:

Related event: Stanford Scholars Discuss CoT Monitoring Limits and AI Safety(2 posts)→

Original post →

More from Safety

Safety channel →