Local GLM 5.3 test: Building a 3D penthouse via BlenderMCP
Fun-Meaning-6474 · reddit · 2026-09-01
The author ran GLM 5.3 and GLM 5.3 Flash (Q4 quantized) locally on multiple RTX PRO 6000 WS GPUs to build a 260 sqm luxury duplex penthouse using BlenderMCP, comparing their performance.
Key Findings:
- Prompting: Vague prompts fail; specific architectural parameters (ceiling height, stair rise, PBR material ranges) are required. The model even added unrequested details like book spines.
- Performance: The full GLM 5.3 spent 21 minutes thinking before placing the first object, taking 40m 43s total and consuming 112K tokens. The Flash version started immediately, taking 38m 52s and using only 36K tokens.
- Accuracy: Flash correctly built the 9x8m double-height void, while the full model built it at 9x4.5m despite reporting it correctly. Flash matched the full model on object count and time with a third of the tokens.
The post includes the exact prompt used with dimensions, materials, and furniture specs.
More from coding & agent
- DIY visual diff tool using GitHub Artifacts and pure JavaScript to save costs — zeeg · 2026-09-01
- Dev shares a dirt-cheap approach to visual diffs — zeeg · 2026-09-01
- Grok Bot automates Shopify updates and supplier coordination — billyjhowell · 2026-09-01
- Grok Bot automates lost deal analysis by mining call and email threads — lennysan · 2026-09-01
- Design pattern: immutable agent artifact revisions behind a stable review URL — RocketSeven · 2026-09-01
- Building a long-term memory benchmark for agents: what to add? — True_Mongoose_7073 · 2026-09-01