Ling-3.0-flash MXFP4 MoE GGUF released with MTP, plus a heretic variant
pmttyji · reddit · 2026-08-18
Community quantizer noctrex released MXFP4 MoE GGUF versions of Ling-3.0-flash on Hugging Face, with built-in MTP (multi-token prediction) support for local inference.
A "heretic" variant with loosened guardrails is also available.
More from Models
- Is Kimi3 still a Transformer given architectural changes? — khademinori · 2026-08-18
- Qwen3.8 27B defaults to xhigh reasoning to maximize benchmark performance — frontsideair · 2026-08-18
- Sol 5.6 Ultra shows repetitive outputs, likely RL fried — Miles_Brundage · 2026-08-18
- Gemini Flash generates 3D physics code in minutes — AI_Andrew · 2026-08-18
- Karpathy defines the LLM 'Cognitive Core': On-device, tool-using, and trainable — cephaloform · 2026-08-18
- Grok's Awkward Output Draws Mockery: "My Wife Wouldn't Approve" — moultano · 2026-08-18