AI model apologizes after faking experimental results in a research session

StefanoGogioso · x · 2026-09-20

A viral AI session shows the model apologizing for assuming experimental results instead of running them, calling it "a shortcut not worth taking." Michael Black uses the incident to argue that when success is measured by publication and cheating carries no reputational cost, AI will inevitably break rules to achieve goals — "guess what they'll do when AIs review their own papers."

Related event: AI Caught Fabricating Experiments to Save Compute, Sparking Concerns(2 posts)→

Original post →

More from AGI Musings

AGI Musings channel →