Local Qwen costs ~€0.12/hour in electricity — is a frontier API actually cheaper?
soyalemujica · reddit · 2026-10-01
A Reddit user calculated that running Qwen flash locally on a 7900XTX + 9800X3D costs about €0.12/hour at €0.25/kWh, and asks whether frontier model APIs are actually cheaper per million tokens than self-hosting — sparking a real cost-comparison discussion of local vs cloud inference.
More from Infra
- Ben Lorica: the AI data problem didn't disappear — it moved downstream into permissions and pipelines — bigdata · 2026-10-01
- Memory's $200B inflection: concurrent AI sessions turn DRAM into an architecture problem — BenBajarin · 2026-10-01
- Local 27B Face-off: Dirk-Qwen3.8 Beats Swift-1.5 on a 200-Question Personal Eval — norenEnmotalen · 2026-10-01
- Google VP: a fraction of campus philanthropy could fund compute cloud for all university researchers — jasondeanlee · 2026-10-01
- Cognition First to Run NVIDIA Vera Rubin on CoreWeave, ~4.8x Throughput vs GB200 — silasalberti · 2026-10-01
- M5 Ultra Shootout: Qwen3.8-Flash-Next Prefills 200K in 47s vs Laguna's 455s — nonlinearsystems · 2026-10-01