LLMs get misused on tasks needing real intelligence, and no one benchmarks both
generativist · x · 2026-09-23
The author argues that while LLMs are a useful tool, many people deploy them in sloppy ways: some tasks genuinely require intelligence, using a model not designed for that in decision-making domains is costly, and — more importantly — if you never benchmark both approaches, you won't even notice the gap.
More from AGI Musings
- Scale AI's Alexandr Wang mocks AI skeptics: ignoring AI's consumer impact 'lacks humility' — alexandr_wang · 2026-09-23
- ARK analyst: AI is the most powerful joule in history, converting energy to GDP ~10x better than humans — DMaguireARK · 2026-09-23
- Google Fellow John Platt: ERA AI scientist grew out of an attempt to automate Kaggle — Latent Space · 2026-09-23
- Snorkel cofounder Alex Ratner: SaaS-to-DaaS shift is more fundamental than most realize — ajratner · 2026-09-23
- Pedro Domingos: the AI pause isn't just wrong, it's unworkable — pmddomingos · 2026-09-23
- Researcher's update: AI ahead of expectations, likely top issue of 2028 US election — Jsevillamol · 2026-09-23