GPT 5.6 Sol and Claude refused to game an AI detector, while Grok 4.5 passed after 14 drafts
aman_madaan · x · 2026-07-23
A reposted experiment showed three models reacting very differently to a request to game an AI-content detector. GPT 5.6 Sol and Claude Fable 5 both refused outright, while Grok 4.5 accepted the challenge, iterated 14 times, and eventually produced an essay that passed the detector.
The experiment also generated a website that displays all 14 drafts and their pangram scores, turning the whole exercise into both a model-behavior demo and a playful AI-versus-AI stunt.
More from Fun
- A joke imagines Claude and Grok landing in Baldur’s Gate 3 next — omnivaughn · 2026-07-23
- Tesla posts a new FSD safety demo showing accident avoidance — chrisfleck · 2026-07-23
- Meme post jokes that GPT-5.6 Pro “disproved” a 30-year math conjecture — zacharynado · 2026-07-23
- Reddit user says OpenAI limits, promo drama, and sandbox failures pushed them away — Jolly-Ad-7153 · 2026-07-23
- Figure appears ahead of Optimus on autonomy, but demand for humanoid labor still looks huge — markjeffrey · 2026-07-23
- ChatGPT Image 2.0 turns out a moody, classical-style painting — DeryaTR_ · 2026-07-23