NVIDIA Ships Qwen 3.8 27B NVFP4 Quantized Model on Hugging Face
TheZachMueller · x · 2026-09-09
NVIDIA has released an NVFP4 quantized version of Qwen 3.8 27B on Hugging Face (nvidia/Qwen3.8-27B-NVFP4), as spotted and shared by TheZachMueller on Twitter.
The model uses the NVFP4 4-bit format optimized for inference on NVIDIA hardware. Its model card includes a chat template supporting image and video inputs, indicating it is a quantized release of a multimodal model, aimed at developers who want to run Qwen efficiently at low precision on NVIDIA GPUs.
Related event: NVIDIA releases official NVFP4 quantized Qwen3.8-27B(2 posts)→
More from Models
- Debate over OpenAI allegedly using user interactions as RL rollouts, calls to publish full solution transcripts — burny_tech · 2026-09-09
- Robot arm self-calibrates with 3 uncalibrated cameras, hits sub-0.2mm accuracy — burny_tech · 2026-09-09
- Ramp data: no-ZDR Fable 5.1 hits 22.5% of enterprise spend, ZDR a hard requirement — zephyr_z9 · 2026-09-09
- ChatGPT web bookmarks fail with React #418 error while in-app works fine — rohanpaul_ai · 2026-09-09
- Scale CEO touts assistant benchmark: Muse scores 9.3, beats Instinct 4-1 on real tasks — alexandr_wang · 2026-09-09
- Podcast: why GPT-6 Astra is so significant and so confounding — The AI Daily Brief · 2026-09-09