Qwen3.8 27B Uncensored GGUF Released with 262k Context and Vision Support
BLUECOW009 · x · 2026-08-23
The Qwen3.8-27B-Unleashed model has been updated with uncensored weights, vision capabilities, and MTP (Multi-Token Prediction) support.
Key Updates:
- Added vision component mmproj-Unleashed-f16.gguf.
- Retained MTP heads (nextn. tensors) for enhanced reasoning.
Performance & Specs:
- VRAM: 13GB (Q3 quantization).
- Speed: 100 tok/s on 1x 4090.
- Context: 262k (needle verified at 250k depth).
- Benchmark: MMLU 82.98% (measured on the 13GB quant).
The release utilizes Unsloth's Dynamic v3 recipe, which allocates bits based on tensor sensitivity (e.g., 8-bit for critical attention gates, 2-bit for less sensitive layers), allowing this 3-bit file to outperform standard 4-bit quants on perplexity.
More from Models
- Flashback: GPT-4 cost $60/M output tokens with 8K context three years ago — gajesh · 2026-08-23
- Open Weights vs Frontier: Just a 3-Point Gap but 1/3 the Cost — MicahBerkley · 2026-08-23
- Ox Alpha Overhyped? Beats GPT-5.6-Luna but Lags Other Frontiers — Al_Grigor · 2026-08-23
- Dev Endorses k3 + NousResearch Harness as Best Combo — markjeffrey · 2026-08-23
- Ornith 1.5 35B Hits 81.8 on GPQA with Thinking Mode, Decodes at 303 tok/s — MikePFrank · 2026-08-23
- Anthropic's Opus 5 and Sonnet 5 flagged as legitimate regressions — bindureddy · 2026-08-23