Qwen3.8-Flash-Next Launches in INT4 and MXFP4; Auto-Round Tool Updated
HaihaoShen · x · 2026-08-31
Qwen3.8-Flash-Next models are now available in INT4 and MXFP4 quantized versions on Hugging Face. A new version of the Auto-Round quantization tool has also been released alongside the models.
More from Models
- Speculation on Permanent Sandbagging by OpenAI and Anthropic — scaling01 · 2026-08-31
- Understanding ChatGPT Work: Simon Willison's Deep Dive — gmays · 2026-08-31
- User Complains GLM 5.3 Flash Overthinks Simple Prompts and Throws Errors — nahmanhuh · 2026-08-31
- Tested: Qwen3.8-Flash-Next Is Faster but Fakes Completion in Hard Tasks — trashacct383 · 2026-08-31
- Google AI Admits to Outputting Racist Content About Latinos — HelpfulQuestions · 2026-08-31
- Taalas demo shows 14,000 tokens/second generation speed — rohanpaul_ai · 2026-08-31