All models get ultrafast: ~300 tps, 8x faster but 6x more expensive

cneuralnetwork · x · 2026-09-30

OpenAI's Ultrafast tier is now available across all models, delivering roughly 8x faster inference at 300 tokens per second — at about 6x the price. Aimed at latency-critical workloads, it forces users to weigh speed against cost.

Related event: OpenAI Launches Ultrafast Tier: 8x Speed at 6x Price Across API, ChatGPT and Codex(13 posts)→

Original post →

More from Models

Models channel →