Qwen3.8 Flash Next NVFP4 Quantized Model Released
RadixArk · hf · 2026-08-31
RadixArk released the Qwen3.8-Flash-Next-NVFP4 model. It supports image-text-to-text tasks, uses FP4 quantization optimized by ModelOpt, and is compatible with the sglang framework.
More from Models
- GLM 5.3 Flash beats Kimi K3 on same coding task at a third of the cost, self-repairs in 10 minutes — HowDevelop · 2026-08-31
- Geology Benchmark: Kimi K3 Leads, GLM 5.3 Flash in Top Tier — teortaxesTex · 2026-08-31
- GPT-5.6 Sol (med) dominates interactive coding agent benchmarks — steipete · 2026-08-31
- Qwen3.8 Flash Next GGUF Version Released — AtomicChat · 2026-08-31
- Single Model Replaces Stack: 61% Cost Cut, Peak Accuracy — DynamicWebPaige · 2026-08-31
- Testing Qwen3.8-27B on Groq: Fast but Lacks Time Awareness — coslinedev · 2026-08-31