Humor may be LLMs' last frontier: who's building the benchmark?
abhishekcode42 · x · 2026-09-13
The author argues LLMs now beat the average human at nearly every word-based skill, leaving humor as one of the last frontiers, and asks who is building a humor benchmark for LLMs — and what other individual skills remain where models still trail average humans.
More from Models
- Forced Model Depreciations May Push Users Toward Open-Weight Alternatives — marlene_zw · 2026-09-13
- Westlake AGI Lab's Code World Model splits world rules (code) from video rendering — jiqizhixin · 2026-09-13
- Abacus AI releases open-weight Smaug Mini 27B, prices Smaug Flash at $0.10/M input — bindureddy · 2026-09-13
- AkbasCore 2.0: damped multi-axis activation steering kernel validated on Qwen2.5-7B — Nearby_Indication474 · 2026-09-13
- Astra ten days later: big CUA/vision gains, no foundational intelligence leap, dev says — tokenbender · 2026-09-13
- Astra beats itself on ARC-AGI-3 with 46% cost cut at max reasoning settings — daniel_mac8 · 2026-09-13