OpenAI Introduces GPT-Live

 

A phone displays ChatGPT with OpenAI in the background
Photo credit: Dreamstime Photos

OpenAI  launched GPT‑Live. The new tool represents a new generation of voice models that make talking with AI feel more like having a real conversation.

 

GPT‑Live is built on a full-duplex architecture, meaning it can listen and speak at the same time. During conversations, GPT‑Live can show it’s paying attention with phrases like “mhmm” or “yeah”, engage in quick back-and-forth, or just stay quiet when you need a moment to think. The result is a voice experience that is refreshingly easy to talk to.

 

For questions that require web search, deeper reasoning or more complex work, it delegates to the latest frontier model behind the scenes and brings the result back into the conversation when it’s ready. While it works, GPT‑Live can keep talking and maintain the flow of conversation. At launch, GPT‑Live will use GPT‑5.5 in the background. As OpenAI releases new frontier models, the company will continuously update GPT‑Live.

 

“These advances power a new ChatGPT Voice experience that is more intelligent and natural to use. Over time, we believe this research will also unlock the ability to use voice for increasingly complex, longer-running and more agentic work,” OpenAI said in its statement.

 

Continuous interaction

OpenAI built GPT‑Live for continuous interaction using a full-duplex architecture.Instead of processing a sequence of separate messages, GPT‑Live continuously processes input while generating output. The model can make interaction decisions many times per second: whether to speak, continue listening, pause, interrupt or invoke a tool.

 

This allows the model to engage in more natural back-and-forth, maintain a better sense of time and even perform live translation.

 

Delegation for deeper work

The company also decoupled GPT‑Live—which handles continuous interaction—from deeper work. When a question requires search, reasoning or more agentic capabilities, GPT‑Live can delegate the task to another model like GPT‑5.5. This allows it to keep the conversation going, even as it handles multiple tasks in the background.

This architectural change also allows GPT‑Live to continuously use the latest models and agents, combining frontier intelligence with natural interaction.

Evaluations

GPT-Live also features human evaluations to measure pleasantness and the flow of conversation. In these head-to-head comparisons, GPT‑Live‑1 and GPT‑Live‑1 mini are strongly preferred over Advanced Voice Mode in matched five to 10 minute conversations that measure overall preference, turn-taking, interruptions, conversational flow and how natural each interaction felt.

A new ChatGPT Voice experience

Each week, more than 150 million people talk to ChatGPT using features like Voice and Dictation. They use it to get hands-free everyday help, to practice languages, tell bedtime stories or just chat during their commute.

 

Now, when you tap the Voice button to talk with ChatGPT, you’ll get an improved experience powered by GPT‑Live—with more natural conversations, smarter answers, better listening and visual responses.

Availability & limitations

GPT‑Live is rolling out now to ChatGPT users globally across iOS, Android and ChatGPT.com⁠. GPT‑Live‑1 will become the default model powering ChatGPT Voice for Go, Plus, and Pro users, and GPT‑Live‑1 mini will become the default for Free users.

 

 

Learn more about GPT-Live here

 

Read more of the latest AI in Eye Care news here

Author

Leave a Reply

Your email address will not be published. Required fields are marked *