Sunday Jul 12

OpenAI Ships Full-Duplex Voice Models

8JUL
TALK + LISTEN

ChatGPT voice can now listen and talk at once. GPT-Live replaces Advanced Voice Mode and handles real interruptions. Harder questions get quietly routed to a bigger model behind the scenes.

GPT-Live is full-duplex, meaning it speaks and listens at the same time.

You can interrupt it mid-sentence, the way you'd interrupt a person.

It drops in small verbal cues, like 'mhmm,' to show it's still listening.

For harder questions, it quietly hands off to GPT-5.5 and brings back the answer.

GPT-Live-1 mini is now the default for free users, GPT-Live-1 for paid tiers.

Video and screen sharing still need the old legacy voice mode for now.

Voice is becoming the interface, not just a chat feature.

full brief & sources

Why this matters

  • Full-duplex voice is a real interaction change, not an incremental voice update.
  • Interruption handling is the detail that makes voice AI feel less robotic.
  • Delegating hard questions to a bigger model behind the scenes is a new architecture pattern worth watching.

🔍 What happened

  • OpenAI released GPT-Live-1 and GPT-Live-1 mini on July 8.
  • Both are full-duplex: they can speak and listen simultaneously, enabling natural interruptions.
  • GPT-Live delegates complex reasoning or search tasks to GPT-5.5 in the background.
  • GPT-Live-1 mini is the new default for Free users; GPT-Live-1 for Go, Plus, and Pro.
  • It's rolling out across iOS, Android, and ChatGPT.com.
  • Video and screen sharing aren't supported yet; those still require the legacy voice mode.

💬 Smart takes

  • OpenAI: GPT-Live can show it's paying attention with small cues, or just stay quiet when you need a moment.
  • SiliconANGLE: the launch lands just ahead of the broader GPT-5.6 release, positioning voice as its own product line.
  • Skeptic: full-duplex demos are easy to show; multi-turn real-world conversations are where the awkward pauses and interruptions usually resurface.

🧭 Where this goes

  1. LikelyOpenAI brings GPT-Live to the API within a few months.
  2. Likelyrival labs ship their own full-duplex voice models within the year.
  3. Possiblevideo and screen sharing merge into GPT-Live by early 2027.
  4. Wild Cardvoice becomes the primary ChatGPT interface for a meaningful share of daily users.

🥄 The Spoon Take

Voice assistants have felt like walkie-talkies for years: talk, wait, listen, repeat. Full-duplex breaks that turn-taking pattern for the first time at this scale. If it holds up in real conversations, voice stops being a feature and starts being the interface.

🤔 Pushback

OpenAI has shipped voice mode updates before that looked great in demos and felt clunky in daily use. The real test is a 20-minute call, not a launch clip.