Writing benchmark: GPT-6.1 Sol regresses 153 Elo below GPT-6 Sol while costing more

OnlyProggingForFun · reddit · 2026-09-30

An internal writing benchmark (181 model configs, 10 script tasks, blind-scored by AI judges from three labs) finds GPT-6.1 Sol is a step back for writing:

Verdict: stay on GPT-6 Sol for writing; if using 6.1 Sol, low is the cheap pick and xhigh the best.

Original post →

More from Models

Models channel →