Testing LLMs with a Churchill Insult Prompt: GPT Outperforms Kimi and Gemini

emollick · x · 2026-07-23

Wharton professor Ethan Mollick tested LLMs by prompting them to generate a "witty Churchill insult." In his latest trial, he found GPT 5.6 Sol Pro to be the winner, with Fable also performing well, while Kimi and Gemini missed the mark by a mile. He previously favored Claude's output.

Original post →

More from Fun

Fun channel →