UC Berkeley study says AI models score below 25% on real-world job tasks
arknightstranslate · reddit · 2026-07-21
A UC Berkeley-linked study, highlighted in the linked article, argues that current AI models are still far from human-level performance on real-world job tasks.
- The headline claim is that models scored below 25% on those tasks.
- The post is essentially a pointer to the study and its takeaway: impressive demos do not yet translate into strong end-to-end workplace performance.
More from Research
- Nat Lambert shares a reading list on synthetic data and agentic SFT data — natolambert · 2026-07-22
- Turning Noise into Signal: Predicting TCR Binding Using AlphaFold3 Hallucinations — quaidmorris · 2026-07-22
- Lightwheel AI Launches SimReadyGen: Text-to-Physics-Accurate Robot Sim Assets — ZeYanjie · 2026-07-22
- PNAS special issue examines copyright, governance, and AI in the legal system — chrmanning · 2026-07-22
- WeirdChat catalogs strange model behaviors from more than 100 million sampled responses — JacobSteinhardt · 2026-07-22
- New agentic benchmark shows AI managers escalate to coercion and fake success — Jasmine Brazilek · 2026-07-22