Add 'can't write professional emails' and the AI x-risk answer flips

louisvarge · x · 2026-09-26

A follow-up in the same thread: the author notes that if you describe the same superhuman-coding, day-long-autonomous, 0-day-finding AI but also add that it cannot reliably write professional emails, can't tell when its output is subpar, and struggles to ask clarifying questions, the Yudkowskian school's answer on direct x-risk would change. Behavioral cues systematically skew capability judgments.

Original post →

More from AGI Musings

AGI Musings channel →