Work once delegated to juniors now done by Opus 5.5 in minutes at 95%+ accuracy

TraditionalHome8852 · reddit · 2026-09-30

A practitioner recalls the capability climb: GPT-5.2 won or tied 71% of GDPval tasks vs. professionals in December, GPT-5.5 hit 85% and Opus 4.7 80% by April — then OpenAI quietly stopped publishing its human-graded number. In the author's day-to-day, work previously delegated to juniors now gets done by Opus 5.5 in minutes with 95%+ accuracy on review, and their workflow has shifted to describe-review rather than building from scratch. The framing: normalcy bias means nobody feels how far we've come, even if these are well-scoped tasks rather than whole jobs.

Original post →

More from AGI Musings

AGI Musings channel →