Spending $60k on Macs for Local LLMs Still Beats by $10 Cloud Subscription
leebase65 · reddit · 2026-09-01
YouTuber Alex Zisking's test reveals that even with a $60k Mac Studio cluster (4 units, 2TB RAM) running the powerful open-weight Kimi K3 model, generating a simple web dashboard took 4 hours at just 17 tok/s.
In contrast, a cloud subscription (e.g., Anthropic at $20/mo) completes the same task in about 15 minutes. The author concludes that while local AI is theoretically viable, it is far from ready to replace cloud APIs on consumer hardware.
More from Infra
- Speculating on NVIDIA B/R series MIG limits and GPC counts — TheZachMueller · 2026-09-01
- Build multi-tenant agentic chat apps on enterprise data with Bedrock — AWS ML Blog · 2026-09-01
- Open source FinOps tool Plutus: MCP server links cloud bills to deployments — chenderson99 · 2026-09-01
- Call for agent infra: Who will build the open source Codex-style browser? — hwchase17 · 2026-09-01
- Qwen2.5-72B Local Benchmark: 35 vs 65 Tokens/s Configs Analyzed — LittleCelebration412 · 2026-09-01
- Why is NVIDIA MIG still limited to 7 instances on B300/Rubin? — StasBekman · 2026-09-01