Hobby Blender benchmark: GPT-6-Astra one-shot scenes strikingly outperform other tested models
Gruku · reddit · 2026-09-08
A developer running a personal Blender benchmark (inspired by MineBench) tested GPT-6-Astra on one-shot Python script scene generation, finding results strikingly better than all other models tried. Blender MCP multi-turn runs looked even more impressive but were too costly to test at scale. The author feels one-shot "make something visually interesting" is hitting its ceiling; next up are game-ready assets and real production tasks. Results and voting at blenderbench.realityreprojector.com.
More from Models
- GPT-6 scores below Kimi K3 on GDPval-AA v2 and neither OpenAI nor Artificial Analysis will explain why — ChrisGPT · 2026-09-08
- IFM Releases K2-Horizon Models, Including a 375B A23B MoE — PhilippeEiffel · 2026-09-08
- davinci-002 shuts down Sept 28: dev recreates the original GPT experience as a farewell — cephaloform · 2026-09-08
- Founder's model ranking: ChatGPT underrated, Gemini 'knocked flat', Claude overhyped — firstadopter · 2026-09-08
- After Days of Coding With Astra, One Dev Calls It Overhyped — Nevetsny · 2026-09-08
- Engineer Joshua Saxe: Astra useful but far from AGI hype — a 'slot machine' AI with no memory of intent — joshua_saxe · 2026-09-08