Wednesday, July 8, 2026

OpenAI releases faster gpt-realtime-2.1 voice models

OpenAI released two new Realtime API models on July 6 - gpt-realtime-2.1 and gpt-realtime-2.1-mini - aimed at building low-latency voice and multimodal agents. The company said improved caching cut 95th-percentile latency by at least 25% across its Realtime voice models, while the full model adds better alphanumeric recognition, silence and noise handling, and interruption behavior. The mini tier brings realtime reasoning and tool use at a lower price point, framing the update as production-focused rather than a new frontier model.

/ Sources

/ About this story

Compiled by Venture Atlas from the sources above, using automated AI-assisted research. This is a summary of reporting published elsewhere, not original reporting - follow the source links for the full account. See our editorial standards.

Something wrong here? Email flightatlas.contact@gmail.com and we will fix it.

/ Related