paraschopra launches crowdsourced study on whether LLMs got funnier at original jokes

paraschopra · x · 2026-09-29

Researcher paraschopra is running a small study probing LLM progress in non-verifiable domains, using humor as the test case: can frontier models write better original jokes than older ones?

To prevent memorization, every joke must include two randomly chosen words, forcing novelty. LLM-as-judge results already show a clear improvement trend over time, and he's recruiting human raters (10 minutes) to ground those judgments; participants get results first by email.

Related event: Crowdsourced Study Tests Whether LLMs Can Write Original Jokes(2 posts)→

Original post →

More from Research

Research channel →