OpenAI self-replicating prompt was misreported, Zvi clarifies it wasn't found in the wild

TheZvi · x · 2026-09-30

David Krueger called out fast-news AI accounts including @TheZvi and @AndrewCurran for multiple misreports of the OpenAI safety story: the original disclosure only said OpenAI "found a self-replicating prompt," not that replication was observed in the wild. TheZvi apologized, saying his initial reaction was too dismissive and the person was communicating something real, and invited criticism of his full writeup.

Related event: Cambridge scholar corrects misinformation around OpenAI self-replicating prompts(3 posts)→

Original post →

More from Safety

Safety channel →