Remodeling LLM intelligence/cost charts: adding GLM-5.3 and local hosting costs

crusaderky · reddit · 2026-08-19

Criticizing Artificial Analysis' (AA) intelligence/cost plot for being misleading (specifically the use of a logarithmic scale for cost), the author created a linear-scale version. The new chart includes GLM-5.3, DeepSeek V4 Flash, and a self-hosted Qwen3.8-27B, with the latter's cost calculated based on electricity prices and RTX 3090 performance. The analysis suggests that for most users, the cost difference between running Qwen locally and using DeepSeek via API is negligible, though local hosting on dedicated high-end hardware (e.g., 128GB RAM) changes the economic equation significantly.

Original post →

More from Infra

Infra channel →