AI Agents Still Can't Conduct Open-Ended Research Due to Cognitive Failures

mikeflache · x · 2026-08-12

A new study shows that while frontier AI agents excel at verifiable tasks like coding, they still fall short in conducting open-ended AI research.

Researchers partnered with authors of two unpublished papers, giving AI agents thousands of dollars in API credits and six days to solve core research questions. The original authors unequivocally rejected the agents' papers. After analyzing over 100 hours of logs, the team found that despite high-level engineering capabilities, agents suffered from key cognitive failures:

Original post →

More from AGI Musings

AGI Musings channel →