Codebase Understanding Test: Big Differences Across Models

emax · x · 2026-07-18

The author reports that on their current codebase, only fable high/xhigh and sol-high/xhigh/max still give fairly accurate codebase reading results.

In contrast, opus 4.8 and gpt-5.5 perform noticeably worse on this codebase, often drawing incorrect conclusions. Overall, sharing practical usability differences in code understanding tasks.

Original post →

More from coding & agent

coding & agent channel →