Qwen3.8 Flash Next GGUF Version Released
AtomicChat · hf · 2026-08-31
AtomicChat released the Qwen3.8-Flash-Next-GGUF model. It is a MoE and multimodal model available in GGUF format for efficient running on llama.cpp, supporting imatrix quantization.
More from Models
- GLM 5.3 Flash beats Kimi K3 on same coding task at a third of the cost, self-repairs in 10 minutes — HowDevelop · 2026-08-31
- Geology Benchmark: Kimi K3 Leads, GLM 5.3 Flash in Top Tier — teortaxesTex · 2026-08-31
- GPT-5.6 Sol (med) dominates interactive coding agent benchmarks — steipete · 2026-08-31
- Qwen3.8 Flash Next NVFP4 Quantized Model Released — RadixArk · 2026-08-31
- Single Model Replaces Stack: 61% Cost Cut, Peak Accuracy — DynamicWebPaige · 2026-08-31
- Testing Qwen3.8-27B on Groq: Fast but Lacks Time Awareness — coslinedev · 2026-08-31