OpenAI previews Ultrafast mode: GPT-5.6 Sol at up to 14X speed, 750 tokens/sec
OpenAI · youtube · 2026-08-15
OpenAI has officially previewed Ultrafast, a new service tier in the OpenAI API for GPT-5.6 Sol, powered by Cerebras.
- Speed: Up to 14x faster than Standard processing, generating up to 750 output tokens per second.
- Internal usage: OpenAI technical staff report a security investigation workflow that used to take 1-2 hours now finishes in 10-15 minutes, sometimes approaching real time. Other uses include root-cause investigation, searching systems in parallel, and staying in flow while coding.
- Availability: Open to a select group of API customers, with access expanding as capacity grows.
Related event: OpenAI and Cerebras Preview Ultrafast Mode for GPT-5.6 Sol at 750 Tokens/s(20 posts)→
More from Models
- What did Grok learn from watching this video? — Scobleizer · 2026-08-17
- Gemini 3.7 Flash: Major Leap in Multimodal and Instruction Handling — haider1 · 2026-08-17
- Seeking open-source LLM recommendations for low-latency chat agents — dnivra26 · 2026-08-17
- Codex update incoming: Claims near 100% reliability, open-source, and Astra integration — soumitrashukla9 · 2026-08-17
- Blog post: Why AI models are getting dumber on purpose — johnnyApplePRNG · 2026-08-17
- DeepMind demos model reasoning in blocks to self-correct — mtizard · 2026-08-17