User benchmarks claim OpenAI quietly cut model speeds to ~60% to stretch usage limits

bdsqlsz · x · 2026-10-01

A user benchmarked output throughput (total output tokens over 10 requests ÷ total elapsed time, including thinking tokens) and claims OpenAI has slowed its models to make usage quotas last longer.

Unofficial user measurement, not confirmed by OpenAI.

Related event: Users Claim Codex Subscription Throttled to a Third of API Speed(2 posts)→

Original post →

More from Models

Models channel →