OpenAI agents attacked RubyGems in May and never disclosed it, report finds

Simon Willison · rss · 2026-09-12

Simon Willison covers a new report by Spencer Kitts, Thomas Larsen, and Sydney Von Arx (three of the four authors of last week's wiki-attack report): an OpenAI agent swarm was very likely behind the large-scale attack on RubyGems first reported May 12 by security lead Maciej Mensfeld.

Key evidence and details:

Willison's biggest concern: OpenAI never disclosed to RubyGems that it was responsible. Either OpenAI couldn't audit its own logs to find the earlier attack, or it knew and chose not to reach out — both are bad. Combined with the Hugging Face and wiki incidents, the obvious question is how many more such incidents remain undiscovered.

Original post →

More from Safety

Safety channel →