Four Intel B70s vs a 128GB MacBook for Local LLMs: A Buyer's Dilemma
Rokett · reddit · 2026-09-21
An M1 Max 64GB user asks for local-LLM hardware advice: current Mac fits large models but runs them too slowly, and replacing a perfectly fine dev machine just for LLM speed feels wasteful.
The choice: an M4/M5 Max 128GB MacBook, or four Intel B70 GPUs at $1,300 each. They note ChatGPT keeps giving conflicting or hallucinated answers and want input from people who actually run these setups.
More from Infra
- One RTX 3090 ran Qwen 27B autonomously for 3 weeks — it shipped working CUDA kernels — skeole · 2026-09-21
- Intel CEO says company can only meet ~50% of server CPU demand — Beth_Kindig · 2026-09-21
- ExLlamaV3 3bpw local Flash benchmarks: 1500 tps prefill on a single RTX 5090 — youcloudsofdoom · 2026-09-21
- SGLang patch brings MCP support: computer use and browser use agents now work — TheZachMueller · 2026-09-21
- Huawei unveils Ascend 960 supernode: 4,096 cards, 8 EFLOPS FP8, NPO optics, Q3 2027 — teortaxesTex · 2026-09-21
- Chinese models' token share on Vercel grew ~5x since June 30 — menhguin · 2026-09-21