Data poisoning: a few hundred crafted docs can backdoor billion-parameter LLMs

ChuckDBrooks · x · 2026-09-23

Cybersecurity expert Chuck Brooks breaks down data poisoning as a top emerging threat to LLMs. Recent research involving Anthropic, the UK AI Security Institute, and the Alan Turing Institute shows roughly a few hundred carefully crafted malicious documents can implant backdoors or alter behavior in models with hundreds of millions to billions of parameters — with the poisoned volume not needing to scale with model size. In his book Inside Cyber, Brooks frames AI as both the strongest weapon in attackers' arsenals and the best defensive tool in what he calls the Acceleration Era.

Original post →

More from Safety

Safety channel →