GPT-5.6 Sol High-Speed Quota Depletion Sparks User Complaints

Recently, the quota depletion rate of the GPT-5.6 Sol model (including High and Ultra modes) has sparked significant user complaints. Multiple users report that previously generous subscription quotas are now being exhausted in a very short amount of time, severely disrupting daily development workflows.

Abnormal Quota Depletion and Degraded Experience

Several users pointed out that the new model consumes quotas at an abnormally fast rate. User @kalmankantaja stated that while using Codex for code refactoring, a 5-hour limit was exhausted in just 15 minutes, with the system warning that less than 10% of the daily quota remained. User @vista8 mentioned that many have been affected by the 5-hour reset rhythm of Codex and Claude Code, noting that GPT-5.6 Sol depletes quotas in just dozens of minutes under High and Ultra modes. User @jdjohnson also reported hitting the 5-hour limit very quickly under their current model setup.

Limited Parallelism and Cost Concerns

The most direct impact of this change is on multi-agent parallel tasks. User @petrusenko_max pointed out that the $200/month Codex plan used to have plenty of headroom, easily handling 10+ parallel agents without hitting the limit. However, under GPT-5.6 Sol, running just 1 or 2 agents often triggers the hard limit. Feedback relayed by @RileyRalmuto corroborated this, noting that running multiple loops simultaneously rarely touched the quota before, making the $200 subscription feel "very worth it"—a stark contrast to the current situation. Furthermore, regarding @sama's claim that this model addresses enterprise cost concerns, @jdjohnson expressed disagreement, stating that the actual experience runs completely counter to those claims.

2026-07-10 ~ 2026-07-11 · 5 related posts

Full story(20 episodes)→