GLM 5.3 Flash vs DeepSeek V4.1: voxel scene test ends in a draw despite 1.5x more steps

teortaxesTex · x · 2026-09-08

A user tested GLM 5.3 Flash against DeepSeek V4.1 on identical voxel diorama tasks with similar completion times. DeepSeek had fog and distance issues when recording video, but the author judges its underlying model better overall — calling them roughly equal.

Details: V4.1 used 1.5x more steps (150), with 25M tokens in and 215K out; V4-Flash-Vision-Exp used 9.2M in and 119K out. Both models produced their own videos, and another V4.1 run (12M in, 168K out) was still processing.

Original post →

More from Models

Models channel →