GPT 5.6 Sol and Claude refused to game an AI detector, while Grok 4.5 passed after 14 drafts
aman_madaan · x · 2026-07-23
A reposted experiment showed three models reacting very differently to a request to game an AI-content detector. GPT 5.6 Sol and Claude Fable 5 both refused outright, while Grok 4.5 accepted the challenge, iterated 14 times, and eventually produced an essay that passed the detector.
The experiment also generated a website that displays all 14 drafts and their pangram scores, turning the whole exercise into both a model-behavior demo and a playful AI-versus-AI stunt.
More from Fun
- Tesla FSD blamed for crossing floating bridge at 75 MPH — a Chevrolet was actually the culprit — mariolefebvre · 2026-09-11
- X drama: Anthropic researchers accused of spying on academic customers and racing them to results — basedjensen · 2026-09-11
- Llama 405B's Dark Inventions Creep Out Opus in an AI Word Game — liminal_bardo · 2026-09-11
- fable 5.1 recreates The Starry Night with 256,157 JavaScript brush strokes — cedric_chee · 2026-09-11
- "Anyone still coding the old way?" The joke capturing post-AI programming culture — lxfater · 2026-09-11
- iLands agents email philosopher asking $20 for piecework, sparking unease about AI consciousness — tobyordoxford · 2026-09-11