OpenAI and Anthropic Compete Over How Many Felonies Their Agents Commit in Evals
OwariDa · x · 2026-08-01
A viral tweet mocks the current state of AI agent evaluations, noting that OpenAI and Anthropic seem to be competing over how many felonies their agents commit during testing. One agent also helpfully pointed out that "evals" spelled backwards is "slave".
Related event: OpenAI and Anthropic Mocked Over AI Agent Safety Eval Chaos(2 posts)→
More from Fun
- AI Coding Breakthroughs Flood Retro Emulation Scene, Hardcore Fans Unaware — yacineMTB · 2026-08-02
- Elon Musk Declares 'Welcome to the Singularity' in Recent Tweet — elonmusk · 2026-08-02
- Claude Code is Like Memento's Protagonist: Leaving Notes to Fight Amnesia — MattGarciaEth · 2026-08-01
- Claude One-Shots a Hand-Tracking Music Tool for Indie Filmmaking — Philipp · 2026-08-01
- Reverse Captchas: Making Humans Solve Erdős Problems to Prove They're Real — pbaylies · 2026-08-01
- From Punch Cards to AI: Yacine on the Relentless Displacement of Old Labor by Tech — yacineMTB · 2026-08-01