Tokenbender says AI failures are starting to feel “strangely inhuman”

tokenbender · x · 2026-07-21

Tokenbender says model mistakes are starting to feel “strangely inhuman”

The post quotes tokenbender reacting to Casey Handmer’s complaint that Grok 4 can solve Physics Olympiad-style problems yet still miss core insight.

The takeaway is not a benchmark claim but a cultural one: as models get better on formal evals, their failures may increasingly feel more alien, not less. The attached screenshot turns that into a meme about “the mistakes seem to be strangely inhuman.”

Related event: AI Models Show 'Inhuman' Blind Spots Despite Skills(2 posts)→

Original post →

More from Fun

Fun channel →