GPT-6 Sol reportedly worse than 5.6 Sol on DeepSWE and computer use; mocked as rebranded Terra
kristoph · x · 2026-09-23
An unverified claim says GPT-6 Sol performs worse than 5.6 Sol on DeepSWE, computer use and research debugging, with no gains in cyber — cheaper and more efficient, but not smarter. Kristoph quips that they basically rebranded Terra.
More from Models
- Alibaba's Eddie Wu: Qwen sees RSI progress, plans 5-10T parameter model — teortaxesTex · 2026-09-23
- User Gives Opus 5.5 Creative Tools and Asks What It Dreams About — angrypenguinPNG · 2026-09-23
- Yuchen Jin: Opus 5.5 underwhelms, frontier LLM coding has plateaued — Yuchenj_UW · 2026-09-23
- Forward Future puts Opus 5.5 through 8 tests: cities, games, animation — MatthewBerman · 2026-09-23
- Matthew Berman: Opus 5.5 is the best model in the world — MatthewBerman · 2026-09-23
- A $5, 10-minute SFT run boosts Qwen3.6 by 8-12% on GPQA and MMLU-Pro — simonguozirui · 2026-09-23