AIToday
Large Language ModelsAudio & SpeechAI Business & IndustryThe Verge AIPublished: Jul 9, 2026, 04:00 JST2 min read

OpenAI upgrades ChatGPT voice with full-duplex model that listens and speaks simultaneously

OpenAI upgrades ChatGPT voice with full-duplex model that listens and speaks simultaneously

Key takeaway

  • OpenAI has launched GPT-Live-1, a new voice model for ChatGPT that can listen and speak simultaneously, eliminating delays and interruptions that plagued the older turn-based system.

  • The full-duplex design enables real-time translation and allows the model to automatically escalate complex queries to more capable text models, making ChatGPT Voice feel more like talking to another person.

3 Key Points

  1. What happened

    OpenAI is rolling out GPT-Live-1, a new voice model for ChatGPT that can listen and speak at the same time, reducing interruptions and enabling real-time translation. The model automatically routes queries to more capable text models like GPT-5.5 when reasoning or web search is needed, and can supplement conversations about weather, stocks, and sports with AI-generated visuals.

  2. Why it matters

    The previous turn-based voice model produced inaccurate answers and struggled with natural conversation flow. The new full-duplex design transforms ChatGPT Voice into a more human-like conversational experience, and the real-time translation feature makes it accessible across language barriers—a practical gain for multilingual users and businesses.

  3. What to watch

    GPT-Live-1 is rolling out across iOS, Android, and web for Go, Plus, and Pro subscribers; a smaller GPT-Live-1 mini model will be the default for free users. OpenAI has added safeguards including crisis helpline support for self-harm conversations and age-appropriate responses for teens.

Ask the AI about this article →

Context & Analysis

OpenAI's shift from turn-based to full-duplex voice architecture addresses a core usability friction in conversational AI. The previous model's reliance on sequential exchanges—where the AI had to await silence before responding—created artificial pauses and decision points that broke the illusion of natural dialogue. By processing input and output streams simultaneously, GPT-Live-1 mimics how human conversation actually works: overlapping speech, real-time feedback, and dynamic topic shifts.

The integration with more capable downstream models (like GPT-5.5) for complex reasoning or web search represents a design choice to optimize latency and accuracy. Rather than forcing all queries through a single voice-specialized model, the system delegates heavy lifting to text-based reasoning engines and returns findings conversationally—a practical solution for queries requiring research or math. The addition of real-time visual supplements (sports scores, weather forecasts) and real-time translation further expands the model's utility beyond pure conversation, potentially making it valuable for business use cases where context-rich communication across languages matters.

OpenAI's explicit mention of safeguards—crisis helpline support, age-appropriate responses for teens, and harm-steering—signals awareness of the mental health litigation risks the company faces. The tiered rollout (full model for paid tiers, a smaller efficient version for free users) reflects both resource constraints and a business incentive to drive subscription conversion.

FAQ

How is GPT-Live-1 different from the old ChatGPT voice model?
GPT-Live-1 is a full-duplex model that can speak and listen at the same time, whereas the previous turn-based model had to wait for you to stop talking before responding. The new model produces more accurate answers, maintains natural conversational flow, and can interrupt itself if you start speaking—features the older model did not support.
Who can use GPT-Live-1 and when is it available?
GPT-Live-1 is rolling out across iOS, Android, and the web for Go, Plus, and Pro subscribers. Free users will receive a smaller, more efficient GPT-Live-1 mini model as the default.
What new capabilities does full-duplex listening enable?
Full-duplex design enables real-time translation while you speak, instead of waiting until you stop talking. You can also ask ChatGPT Voice to stop talking until called on, and it will acknowledge it is listening with phrases like 'mhmm,' 'yeah,' or 'got it.'

Get the latest Large Language Models news every morning

For example, today's edition would include:

  • Nvidia revives Rubin CPX chip with major redesignYahoo Finance AI · 2h ago
  • AI advice followed by 79%, but well-being unchangedITmedia AI+ · 5h ago
  • Enterprises face agent governance gapSiliconANGLE AI · 8h ago

AI-summarized, only the topics you pick — one digest a day via Email, Slack, or Discord.

Free · takes 30 seconds · unsubscribe anytimeWhat is AIToday? →

Ask AI

Ask AI anything about this article. Q&As are published on this page for other readers too.

Related Articles

Next articleOpenAI releases full-duplex voice models for longer, more natural ChatGPT conversations