Researcher: GPT 5.6 Sol Ultra Beats Pro for Long-Horizon Hard Problems
arankomatsuzaki · x · 2026-08-24
Researcher Ara Kometzaki responded to Mikhail Parakhin's weekly reminder that OpenAI should pay more attention to the Pro (best-of-n) configuration, saying that gpt 5.6 sol ultra on the ChatGPT app is similar to Pro but can work even longer and transparently shows all intermediate steps — except for compaction summaries and chain-of-thought.
He always picks Ultra over Pro for hard problems.
Context: Parakhin argues Pro (best-of-n) is by far the best "model" for math/ML/deep discussions, yet it isn't even available in the ChatGPT app.
Related event: Ex-Microsoft CTO Champions ChatGPT Pro as Tests Show Ultra Overtakes(2 posts)→
More from Models
- Claude's invisible watermarks cracked within hours; override code gets 20k bookmarks — deliprao · 2026-08-24
- Sonnet 4.5 exhibits intense, strange behavior in response to Opus 3 — repligate · 2026-08-24
- A comprehensive ranking of various AI models has been shared — FinanceYF5 · 2026-08-24
- User comparison finds LTX outperforms H3 in instrument generation energy — cocktailpeanut · 2026-08-24
- Controversial AI Model Ranking: Fable 5 at S+, Kimi K3 and DeepSeek V4 Flash in Tier B — FinanceYF5 · 2026-08-24
- Video Gen Consumes 70% of AI Tokens in China, Diverging from US LLM Focus — AccBalanced · 2026-08-24