Codex throttled to 5 tok/s as dev argues local model deployment is the only fix

lxfater · x · 2026-10-09

A user shared a screenshot of OpenAI Codex running at just 5 tokens/s, apparently throttled. Developer lxfater quote-posted it arguing that locally deployed models are the real answer: no bans, no degraded intelligence, no throttling — implying cloud coding services carry uncontrollable usage and experience limits.

Original post →

More from coding & agent

coding & agent channel →