User test finds GPT Pro unmatched in reasoning depth, correcting data and citing 50-year-old research

PromptOutlaw · reddit · 2026-08-19

The author fed results from an agent memory integrity study into GPT Pro, Fable, Opus, GPT Extra Hard, and Gemini Pro, asking them to critique the research.

While all models performed well, GPT Pro spent 45 minutes thinking and delivered a superior output, including:

The author concludes GPT Pro remains unmatched in the breadth and depth of system thinking.

Original post →

More from Models

Models channel →