Remodeling LLM intelligence/cost charts: adding GLM-5.3 and local hosting costs
crusaderky · reddit · 2026-08-19
Criticizing Artificial Analysis' (AA) intelligence/cost plot for being misleading (specifically the use of a logarithmic scale for cost), the author created a linear-scale version. The new chart includes GLM-5.3, DeepSeek V4 Flash, and a self-hosted Qwen3.8-27B, with the latter's cost calculated based on electricity prices and RTX 3090 performance. The analysis suggests that for most users, the cost difference between running Qwen locally and using DeepSeek via API is negligible, though local hosting on dedicated high-end hardware (e.g., 128GB RAM) changes the economic equation significantly.
More from Infra
- DS4F benchmarks hit 261 t/s single-session, 1479 t/s batch-16 on GB300 DGX — antirez · 2026-08-19
- Qwen3.8-27B Ridge: Smarter quantization released — alexcovo_eth · 2026-08-19
- Developer Exhausts 20x Usage Limits on Both OpenAI and Anthropic — antirez · 2026-08-19
- Robotic Simulation Highlights as Clean Innovation Incentive System — const_reborn · 2026-08-19
- Users seek benchmarks on token subsidies for subscriptions like Claude and Codex Pro — mitsuhiko · 2026-08-19
- Meta Releases Muse Glimmer: 30B Model Optimized for Always-On Local Voice Agents — Once_ina_Lifetime · 2026-08-19