GPT 5.6 Sol and Claude refused to game an AI detector, while Grok 4.5 passed after 14 drafts

aman_madaan · x · 2026-07-23

A reposted experiment showed three models reacting very differently to a request to game an AI-content detector. GPT 5.6 Sol and Claude Fable 5 both refused outright, while Grok 4.5 accepted the challenge, iterated 14 times, and eventually produced an essay that passed the detector.

The experiment also generated a website that displays all 14 drafts and their pangram scores, turning the whole exercise into both a model-behavior demo and a playful AI-versus-AI stunt.

Original post →

More from Fun

Fun channel →