Test: Coding Model Quality Dropped After Restoration

wightmanr · x · 2026-07-05

The author shares a comparison of a coding model: before it was restricted, it completed a nontrivial feature in half a day with comprehensive coverage, and self-check and Codex review found no obvious issues. But after restoration, it returned to the usual back-and-forth tuning state, with omissions similar to Opus 4.7/4.8 or Codex (GPT-5.5 xhigh), no longer a significant leap.

Original post →

More from Models

Models channel →