- OpenAI is debuting GPT-Live, an upgraded chat mode designed to be a much better listener with full-duplex capability.
- Because it continuously processes input while generating output, GPT-Live is better at deciding when to talk, interrupt, or listen.
- GPT-Live can deploy agents to conduct web searches while talking, eliminating pauses.
- GPT-Live is rolling out to all ChatGPT users including free users starting today.
OpenAI's new GPT-Live voice models for ChatGPT introduce a significant upgrade in conversational AI, featuring full-duplex audio that allows the model to listen and talk simultaneously. This enhancement is designed to create a more natural interaction, akin to a standard phone call, where interruptions and overlapping speech are possible.12
The new system, which includes two versions—GPT-Live-1 and GPT-Live-1 mini—is rolling out to all ChatGPT users, including those on free accounts. This means that users can now engage with the AI in a more fluid manner, as it can continuously process input while generating output. This capability allows GPT-Live to better determine when to speak, interrupt, or remain silent, enhancing the overall user experience.5
Additionally, GPT-Live can deploy agents to conduct web searches while conversing, eliminating the pauses that previously occurred when the AI needed to look up information. This feature aims to make interactions more seamless and engaging, as users will no longer experience delays during conversations.
“Now OpenAI is debuting an upgraded chat mode that’s also designed to be a much better listener,” a representative stated, highlighting the focus on improving user engagement. The rollout of GPT-Live marks a significant step forward in making ChatGPT a more interactive and responsive tool for users.
Overall, the introduction of GPT-Live represents a major advancement in AI communication technology, making ChatGPT a more valuable conversational partner.
“GPT-Live is rolling out to all ChatGPT users, including free users, starting today. The model continuously processes input while generating output, and can deploy agents to conduct web searches simultaneously — eliminating the 'searching the web' pause that occurred in the old voice mode.”

