OpenAI Calls Realtime Voice Model 2.1 Its Best Yet

pbbakkum · x · 2026-07-07

OpenAI's Patrick Bakkum stated that gpt-realtime-2.1 is currently their best speech-to-speech model, while the 2.1-mini is the first small audio model released in quite some time, encouraging users to try them out and provide feedback.

Related event: OpenAI Launches GPT-Realtime-2.1 Audio Models(4 posts)→

Original post →

More from Multimodal

Multimodal channel →