DFlash 2 available for Qwen 3.8 27B and Muse Glimmer
rerri · reddit · 2026-08-19
DFlash 2 is now available for Qwen 3.8 27B and Muse Glimmer models. GGUF quantizations have been released alongside a llama.cpp PR to support the functionality.
More from Infra
- NVIDIA publishes multi-GPU method for massive-scale UMAP in minutes — leland_mcinnes · 2026-08-19
- Local inference economics: $60/month power bill for slow speeds — Thin_Pollution8843 · 2026-08-19
- Australia offers free midday power, challenging space datacenter economics — aronchick · 2026-08-19
- DFlash2 on Qwen3.8 27B hits ~200tk/s for code, requires more VRAM — Hefty_Wolverine_553 · 2026-08-19
- ZML runtime now supports 9 hardware platforms including NVIDIA, AMD, and MooreThreads — ylecun · 2026-08-19
- Cerebras holds first conference as public co, chips power OpenAI — Scobleizer · 2026-08-19