LLM inference 101: bandwidth ÷ weight bytes gives your throughput ceiling before any code

abhijithneil · x · 2026-09-04

Engineer abhijithneil started a thread arguing LLM inference is where AI engineers can now get the 'knowing the theory' dopamine rush — and it's a job AI can't easily automate away.

Key points (including replies):

Original post →

More from Infra

Infra channel →