Dev claims benchmarks are 'absolutely meaningless' — models only differ by vibe

gnukeith · x · 2026-09-22

Developer gnukeith argues that benchmarks are "absolutely meaningless," that claims about the scale of intelligence are hype and marketing, and that the only real difference between models is the "vibe" — the subjective feel of using them. The post taps into growing skepticism in the AI community that public leaderboards diverge from real-world experience.

Related event: Dev claims benchmarks are meaningless, only 'vibes' separate models(2 posts)→

Original post →

More from Models

Models channel →