Tired of Cheesed Benchmarks: Redditors Ask Where to Find Trustworthy LLM Comparisons

Sarlo10 · reddit · 2026-09-15

A Reddit user vents that content creators hype every new model as AGI while benchmarks keep getting cheesed, and asking AI itself just surfaces unreliable sources. They ask where to find trustworthy info on how LLMs actually compare in real usage — including whether Chinese models are genuinely that good or just prompt-dependent, unlike Claude which excels even with bad prompts. The thread probes the reliability of benchmark-driven coverage versus real user experience.

Original post →

More from Models

Models channel →