antirez releases DeepSeek V4 Flash GGUF quantizations, from 2-bit to 4-bit, for local deployment
antirez · x · 2026-08-01
Redis creator antirez has released GGUF quantizations of DeepSeek V4 Flash on Hugging Face, including IQ2XXS, Q2K, Q4K, mixed-precision, and MTP versions, totaling 2.79 TB. The files have passed smoke tests but not full testing, with antirez planning to test them tomorrow.
More from Models
- Claude Dominates Web Dev AI Leaderboard, Kimi K3 Secures Top 3 — arena · 2026-08-01
- Researcher Praises Open Source AI: Gap with Frontier Closing Rapidly — Xianbao_QIAN · 2026-08-01
- OpenAI Turns Reasoning into a Budget Line: Return on Cognitive Spend — krishnan · 2026-08-01
- OpenAI demos unreleased 'Astra' model to policymakers, source says — CremeSubject7594 · 2026-08-01
- OpenAI reportedly preparing new model family 'Astra' for multi-agent long-horizon tasks — thesaraharminta · 2026-08-01
- DeepSeek v4 Flash GGUF quant released for DS4 engine, doubles speed to 30+ tok/s — returnity · 2026-08-01