astra claims breakthrough: generating videos by rendering Blender 3D scenes instead of pixel-by-pixel diffusion

IndraVahan · x · 2026-09-06

IndraVahan says astra has made a significant leap toward a new video generation paradigm: instead of diffusion models building frames pixel-by-pixel, it constructs Blender 3D scenes that are rendered as video.

The argument: no matter how good a diffusion model gets, pixel/frame generation always leaves traces of being AI-generated, while 3D models are more solid and offer far greater control over scenes and every asset in them. This reframes "multimodality" as actually building assets to render frames.

The author believes this route will advance not just video generation but world models, gaming and animation, with platforms like Meshy playing a big role.

Original post →

More from Multimodal

Multimodal channel →