Stanford professor challenges DeepMind biosecurity watermark paper as trivially removable

anshulkundaje · x · 2026-10-04

Professor Anshul Kundaje publicly criticized Pushmeet Kohli's DeepMind biosecurity watermark research. While the paper's discussion acknowledges caveats, he argues evidence from Rohit shows that for the reported results it is trivial to eliminate the watermark — a valid critique he asks the authors to address directly. He finds the algorithm interesting as a pilot study with potential biosecurity applications, but says the paper fails to substantiate its claims.

Related event: Stanford Professor Challenges DeepMind's AI Biosecurity Watermark Paper(4 posts)→

Original post →

More from Safety

Safety channel →