NVIDIA's NVFP4-quantized Qwen3.8-Flash-Next image-text-to-text model trends on Hugging Face
nvidia · hf · 2026-09-05
NVIDIA released Qwen3.8-Flash-Next-NVFP4 on Hugging Face, where it is trending. The model is an FP4-quantized build of Qwen3.8 produced with NVIDIA's Model Optimizer (ModelOpt), distributed in safetensors format with an image-text-to-text pipeline, targeting low-precision multimodal inference deployment.
Related event: NVIDIA Releases NVFP4 Quantized Qwen3.8-Flash-Next(2 posts)→
More from Models
- GPT-6 availability is a mess: Pro tier in Chat, all efforts in Work, absent in Codex — justalexoki · 2026-09-05
- Asking Astra to generate an animation with both time and space symmetries — yaroslavvb · 2026-09-05
- 3D artist: GPT-6 assembles and animates a whole car from primitives in one prompt — petewoodbridge · 2026-09-05
- Same insurance table query: Ministral 14B and Qwen3.8-27B nail it, Gemma 4 31B hallucinates — andrejusb · 2026-09-05
- Astra Computer Use Takes Five Minutes Per Step in Real Testing — bubu19999 · 2026-09-05
- GPT 6 Astra day-one impressions: fast, good with skills, solid bug-finding — cneuralnetwork · 2026-09-05