GPT-6 Sol scales almost linearly with effort level, still trails Opus 5.5 in tests

PawelHuryn · x · 2026-09-23

Developer Pawel Huryn shares live-benchmark comparisons across GPT-6 Sol effort levels (n=3 for max, n=1 otherwise), showing scores rise in an almost straight line with effort. Earlier data points: GPT-6 Sol (max) ≈ GPT-5.6 Sol (medium), but still below Opus 5.5 (medium). GPT-6 Luna and Terra are queued up next for testing.

Related event: GPT-6 Sol Scales Near-Linearly With Effort but Lags in Bug Hunt(2 posts)→

Original post →

More from Models

Models channel →