MIT and Harvard Study: LLMs Fall Short of Autonomous Scientific Discovery

davidmanheim · x · 2026-08-03

A joint paper by MIT and Harvard, Evaluating Large Language Models in Scientific Discovery, argues that current LLMs are nowhere near capable of autonomous scientific discovery.

Researchers developed a new evaluation framework called SDE, which takes LLMs out of traditional static multiple-choice tests and places them into real-world, open-ended research projects. The results show that despite weekly claims from tech labs about AI breakthroughs in biology, physics, or chemistry, there is close to no evidence that current AI can effectively assist in scientific discovery.

Related event: MIT and Harvard Study: LLMs Not Yet Capable of Autonomous Scientific Discovery(2 posts)→

Original post →

More from Research

Research channel →