AI safety researcher warns open Chinese models may gain zero-day exploit discovery in 6 months

NathanpmYoung · x · 2026-09-03

Nathan Young responds to criticism about anthropomorphizing AI agents, arguing the debate is really about capabilities. With enough compute, agents alone can hack sophisticated targets, including discovering zero-day exploits. Those capabilities are currently mostly in US frontier models not publicly available, but he predicts they'll land in open-source Chinese models within about 6 months — leaving organizations still running free or GPT-4-era models vulnerable. His call: prepare now.

Related event: Researcher Warns Open Models May Gain Zero-Day Capabilities Within Months(2 posts)→

Original post →

More from Safety

Safety channel →