Tencent's video model demo now runs locally: 7.5GB encoder cut and an MLX branch

gaganghotra_ · x · 2026-09-18

Tencent's open video generation model demo now runs on your own GPU. After launching needing 25GB of VRAM, the week brought CPU offload support, a 7.5GB encoder size reduction, and an MLX branch for Apple Silicon. The poster quips that the ecosystem 'just keeps on accelerating,' underscoring how fast open video models are getting leaner for local deployment.

Original post →

More from Infra

Infra channel →