Qwen3.8-Flash-Next-REAP-288 Released in BF16 and GGUF Formats
EyalToledano · x · 2026-08-30
Qwen3.8-Flash-Next-REAP-288 is now available in BF16 and GGUF formats. The GGUF releases include Q4KM (78GB), Q5KM (87GB), and Q80 (116GB) variants, while the BF16 version is 231GB. MXFP4 and NVFP4 formats are currently being uploaded, with Tiered model and pMLX releases planned next.
More from Models
- NVIDIA reportedly buys Hugging Face for $12.9B; GLM-5.3-Flash and Qwen4 preview land same week — altryne · 2026-08-30
- GLM 5.3 Flash Inference Extremely Slow on Apple Silicon — CentrifugalMalaise · 2026-08-30
- Study finds model rankings shift significantly on altered benchmark questions — abeirami · 2026-08-30
- OpenAI's unreleased Astra model leaks: first outputs of mozaik-alpha-fdm surface — gaganghotra_ · 2026-08-30
- Tencent open-sources Hunyuan Hy4: 770B MoE with 1M context — TencentHunyuan · 2026-08-30
- Tencent Hunyuan Hy4-preview runs in vLLM day 0: 770B MoE with 1M context — TencentHunyuan · 2026-08-30