After NVIDIA's llama.cpp acquisition, are used V100s still a safe cheap-VRAM bet?
OnlineParacosm · reddit · 2026-08-28
A Reddit user worries NVIDIA's llama.cpp acquisition makes old cards a risky investment: the only current cheap path to VRAM and performance seems to be the v100 (now on driver version 580), and if NVIDIA drops support it would be a huge blow to local inference.
They also note CoreWeave's recent massive backstop that kept A100s off the market — previously the easiest way to get fast 80GB in a workstation form factor for under $5k — making the sxm2 Volta route questionable. They ask the community what everyone's plan is.
More from Infra
- Optimizing Minimax H3 Inference Speeds on Consumer Hardware — Ambitious_Fold_2874 · 2026-08-28
- Acorn: AI Chief of Staff running on your own server — MatthewChang · 2026-08-28
- Trade unions push back against data center opposition — MatthewBerman · 2026-08-28
- Implementing Automatic Model Routing Based on Intent — chadwell · 2026-08-28
- UK Labour rejects Green party call to halt AI datacentre construction — nordicinst · 2026-08-28
- llama.cpp mmap fits Qwen3.8-Flash-Next in 16G+64G RAM at 26t/s — q8019222 · 2026-08-28