Qwen3.8 Multimodal Model Released in GGUF Format for Local Inference
HauhauCS · hf · 2026-08-18
HauhauCS released a GGUF version of the Qwen3.8-27B model, a multimodal vision model supporting speculative decoding and FastMTP. The GGUF format facilitates local execution on consumer hardware, optimizing for efficient inference in resource-constrained environments.
Related event: Qwen3.8-27B Multimodal Model Released in GGUF Format for Local Inference(2 posts)→
More from Models
- Karpathy defines the LLM 'Cognitive Core': On-device, tool-using, and trainable — cephaloform · 2026-08-18
- Grok's Awkward Output Draws Mockery: "My Wife Wouldn't Approve" — moultano · 2026-08-18
- Sakana Namazu model now available on Vercel AI Gateway — SakanaAILabs · 2026-08-18
- Claude Fails Miserably at Reading Markdown Files — yoobinray · 2026-08-18
- Ling-3.0-flash MXFP4 MoE GGUF released with MTP, plus a heretic variant — pmttyji · 2026-08-18
- Does High Concurrency Make MoE Serving Load Nearly All Weights Per Token? — LocalLLaMa_reader · 2026-08-18