Wharton Report: Chain-of-Thought Prompting Shows Diminishing Returns in Modern Models

emollick · x · 2026-08-02

A new technical report from Wharton Generative AI Labs reveals that the classic "think step by step" (Chain-of-Thought) prompting is losing effectiveness on modern AI models.

By testing each question 25 times per condition, the research exposes inconsistencies masked by traditional one-time tests, challenging the assumption that CoT is universally beneficial.

Original post →

More from Research

Research channel →