GPT-5.6 Sol Scores Higher on Sol Benchmark but Costs More
ENT_Alam · reddit · 2026-07-15
A Reddit post compared the 3D block-building performance of GPT-5.5 Pro and GPT-5.6 Sol on MineBench.
Results Overview
- Average reasoning time: 25m16s
- Total cost for 15 builds: $710.82, averaging $47.39 per build
- This is one of the most expensive models the author has tested to date, costing over 3 times more than the previous GPT-5.5 Pro version
Experience Conclusion
- GPT-5.6 Sol generates more detailed and creative structures
- It is less conservative, occasionally adding extra details like scarecrows or drying racks
- Overall better sense of scale and structural integrity
- The tradeoff is that the output JSON is typically larger, leading to significantly higher running costs
The author's TL;DR is: Highly capable model, but extremely expensive.
Related event: GPT-5.6 Sol Benchmarked on MineBench(2 posts)→
More from Models
- Same Echo Maze prompt, three frontier models: all passed visually but shipped the same hidden bug — eyishazyer · 2026-09-11
- Benchmark scores drop from 89% to 19% on new evals — how benchmaxxing breaks leaderboard trust — airesearch12 · 2026-09-11
- ChatGPT tells user their question is too hard and to 'accept dumber answers' — phido3000 · 2026-09-11
- Claude is no longer available for minors as Anthropic rolls out age assurance — Muhammad523 · 2026-09-11
- Developer Building a Unified Leaderboard of All Model Benchmark Scores — airesearch12 · 2026-09-11
- Rumor claims Kimi faked performance by serving Claude; DeepSeek new model surprises in evals — realsohamparekh · 2026-09-11