GLM5.3-Flash nvfp4 quantized build for DGX Spark trends on Hugging Face

autotrust · hf · 2026-10-08

A quantized build of GLM5.3-Flash (E224), packaged as autotrust/GLM5.3-Flash-E224-DGX-Spark, is trending on Hugging Face. The image-text-to-text MoE model uses nvfp4 quantization via modelopt, ships in safetensors, is optimized for vLLM, and targets NVIDIA's DGX Spark (GB10/Blackwell) for local on-device deployment.

Original post →

More from Models

Models channel →