NVIDIA ships NVFP4-quantized Qwen3.8-27B, trending on Hugging Face
nvidia · hf · 2026-09-11
NVIDIA released an NVFP4-quantized version of Qwen3.8-27B on Hugging Face, now trending on the model leaderboard. It was produced with the Model Optimizer (ModelOpt) pipeline, ships as safetensors, targets FP4 precision with FP8-related tags for inference cost optimization, and works out of the box with the text-generation pipeline.
Related event: Nvidia Releases NVFP4 Quantized Qwen3.8-27B, Tops Hugging Face Trending(3 posts)→
More from Infra
- 1:26 continuous aerial AI video made entirely on a Mac with MiniMax H3 — cocktailpeanut · 2026-09-11
- KV cache gets QAT too: why this model beats others at fp4 KV cache — stochasticchasm · 2026-09-11
- Commentary: Anthropic loads shift to Google plus AWS slice, OpenAI doubles down on Azure — ericwdolan · 2026-09-11
- Nvidia claims Vera Rubin delivers 50X throughput per MW and 35X lower token cost vs Blackwell Ultra — Beth_Kindig · 2026-09-11
- Epoch AI: GPT long-context latency scales quadratically, matching price jumps — Jsevillamol · 2026-09-11
- RTX 3090 mini-bench: ninfer cuts TTFT from 3.4s to 29ms, prompt processing ~76x faster — milkipedia · 2026-09-11