Mapping LLM Size Tiers to Hardware Specs

DeepOrangeSky · reddit · 2026-07-09

The post asks what hardware constraints dictate common local LLM size tiers (like 30B, 70B, 120B, 230B). Is it based on professional GPUs, consumer GPU VRAM, or a mixed consideration post-quantization?

The author specifically wants to understand why these common sizes have formed fixed "tiers" and what server or desktop deployment conditions these tiers correspond to.

Original post →

More from Infra

Infra channel →