ChatGPT Voice Can Now Listen While It Talks


TL;DR

  • New Voice Defaults: OpenAI is making GPT-Live-1 the default ChatGPT Voice model for paid consumer users, while Free users get GPT-Live-1 mini.
  • Why It Matters: GPT-Live uses a full-duplex design, so ChatGPT can process speech while speaking instead of waiting for a clean turn to end.
  • Consumer Rollout: The update is rolling out on chatgpt.com and the iOS and Android apps in supported regions, but Business, Enterprise, Edu, video, and screen-sharing support are not included at launch.
  • Open Questions: OpenAI says real-time safeguards can steer, interrupt, or end risky voice conversations, but reliability across noisy, multilingual, and emotionally sensitive real-world use remains the key test.

OpenAI is replacing the default models behind ChatGPT Voice with GPT-Live-1 and GPT-Live-1 mini, giving consumer users a voice system designed to handle interruptions, pauses, and spoken follow-ups more naturally.

The rollout splits access by plan. GPT-Live-1 powers ChatGPT Voice for paid Go, Plus, and Pro users, while Free users get the smaller GPT-Live-1 mini model. OpenAI says the update is rolling out across chatgpt.com and the ChatGPT apps for iOS and Android in supported regions.

The change matters because voice assistants often struggle with timing. Older systems typically wait for the user to stop talking before answering, which can make a brief pause, background noise, or mid-sentence correction feel like the end of a turn. GPT-Live is meant to reduce that friction by keeping the model engaged while a conversation is still unfolding.

OpenAI says more than 150 million people use ChatGPT voice features such as Voice and Dictation each week. For those users, the update is less about a new model name than a change in behavior: ChatGPT should be able to wait while someone thinks, stop when interrupted, and keep track of a spoken exchange without forcing every interaction into rigid turns.

 

What GPT-Live Changes in ChatGPT Voice

GPT-Live uses a full-duplex architecture, meaning it can process incoming speech while producing outgoing speech. Instead of treating conversation as a sequence of separate messages, the model repeatedly decides whether to speak, keep listening, pause, interrupt, or call a tool.

The practical result: ChatGPT Voice should behave less like a walkie-talkie and more like a conversation partner that can hear what is happening while it is talking.