Microsoft shows DeepSeek V4 Flash, a 284B model, running locally on Windows at 1.6-bit quantization
rohanpaul_ai · x · 2026-10-08
Microsoft showcased DeepSeek V4 Flash, a 284B-parameter open-weight model that runs locally on a Windows PC once quantized to 1.6 bits (60GB of memory). The new Surface Laptop Ultra, built on Nvidia's RTX Spark chip with up to 128GB unified memory, has room for it in higher-memory configs. Full video on the Windows YouTube channel.
More from Infra
- Reddit asks: have AI scaling laws hit their limit, or is the compute buildout just starting? — StupidDialUp · 2026-10-08
- TRIAGE stabilizes native NVFP4 RL training, hits full-precision quality at 2.3x throughput — InfiX-ai · 2026-10-08
- WSL Containers Now Generally Available: Run Linux Containers Natively on Windows — pavandavuluri · 2026-10-08
- ai& Says It's Japan's Largest Dedicated Inference Provider, Teases Post-Training Offerings — DavidBennett__ · 2026-10-08
- Microsoft unveils Surface Laptop Ultra with Nvidia RTX Spark SoC from $2,599 — Ars Technica AI · 2026-10-08
- DeepSeek V4.1 shrinks cache 437x, Flash beats V4-Pro 39 vs 36 at half the cost — DeepLearningAI · 2026-10-08