Zitronbench: user challenges critic Ed Zitron to name a task SOTA LLMs fail
iruletheworldmo · x · 2026-09-06
An X user has publicly challenged AI critic Ed Zitron to name a simple task that current SOTA large language models get wrong. The challenge, dubbed zitronbench, invites anyone in the replies to one-shot Zitron's task with proof it was LLM-only. It's a public wager against the recurring claim that LLMs fail at even the most basic tasks.
More from Fun
- Turning AI-generated pixel sprites into animations with Sprite Fusion — HugoDuprez · 2026-09-07
- 2,056 agents with independent weights battle for #1 in a single 24-hour environment — jsuarez · 2026-09-07
- mitsuhiko's 'slop factory' autonomously decides to implement braces and lexical scoping — mitsuhiko · 2026-09-07
- Grok, ChatGPT and Gemini All Failed at Making a Collage of His Book Covers — pickover · 2026-09-07
- The brutal truth of vibe coding: fix one bug, three more appear — ayushtweetshere · 2026-09-07
- When you ask AI to change a CSS class and it suggests rm -rf node_modules — ThePeterMick · 2026-09-07