Testing LLMs with a Churchill Insult Prompt: GPT Outperforms Kimi and Gemini
emollick · x · 2026-07-23
Wharton professor Ethan Mollick tested LLMs by prompting them to generate a "witty Churchill insult." In his latest trial, he found GPT 5.6 Sol Pro to be the winner, with Fable also performing well, while Kimi and Gemini missed the mark by a mile. He previously favored Claude's output.
Related event: Wharton Professor Tests LLMs on Churchill-Style Insults(2 posts)→
More from Fun
- AI safety isn't a coordinated cabal: half the field has posted their life stories on LessWrong — ShakeelHashim · 2026-09-11
- Kid Coins "Princessmaxxing" After Subway Chat About Same-Sex Wedding Attire — anderssandberg · 2026-09-11
- 'AGI is here' vs reality: AI labs still ship some of the jankiest desktop apps ever — MilesCranmer · 2026-09-11
- 'The revolution will have a token limit': one-liner on context window limits — AIandDesign · 2026-09-11
- One-liner echoing the nostalgia: missing human craft, writing, and technical debates — vboykis · 2026-09-11
- Fake Zen saying about bullying X gurus who sell courses and coaching goes viral — DionysianAgent · 2026-09-11