Comparison: Claude Beats GPT-5.5 at Understanding Visual Gags

Sauers_ · x · 2026-07-07

@Sauers used a humorous scenario—a grandson holding a book upside down to pretend he is reading in front of his grandma—to comparatively test two models. Claude correctly identified that the humor stems from the grandson thinking he fooled his grandma while the audience knows it's ridiculous. In contrast, GPT-5.5 incorrectly cited the "upside-down book" as evidence for the audience catching his fake reading (in reality, the audience already knew he was pretending from the context, regardless of the book's orientation).

This comparison highlights the behavioral differences between the two models in multimodal humor comprehension and reasoning.

Related event: GPT-5.5 Stumbles on Basic Reading Tests Against Claude(3 posts)→

Original post →

More from Models

Models channel →