GPT Live | ChatGPT Adds Full-Duplex Voice for Creative Workflows

OpenAI has introduced GPT Live, a new generation of voice models designed to make conversations with ChatGPT feel more continuous and responsive. The system can listen while speaking, recognize when someone needs time to think, handle interruptions more naturally, and delegate complex work without ending the conversation.


GPT Live full-duplex voice experience in ChatGPT

{getToc} $title={Table of Contents}

A more continuous way to work with ChatGPT Voice


GPT Live replaces the rigid turn-by-turn rhythm that has shaped many voice assistants. Its full-duplex architecture continuously processes incoming audio while producing a response, allowing the model to decide whether it should speak, keep listening, pause, respond to an interruption, or use another tool.


For creators, this can make spoken brainstorming, project planning, language practice, and early creative review feel less like recording a sequence of separate prompts. ChatGPT can acknowledge that it is listening, wait while an idea is being developed, or adjust its pace when asked without requiring the user to restart the discussion.



GPT Live can listen and speak at the same time


Previous real-time voice models processed audio directly but still separated the conversation into distinct turns. A pause or unexpected background sound could be interpreted as the end of a prompt, causing ChatGPT to respond before the user had finished explaining an idea.


GPT Live keeps listening while it speaks, making interruptions and quick exchanges easier to manage. It can also stay quiet when asked, focus more effectively on the user's voice in noisy environments, perform live translation, and respond with short acknowledgements that make it clear the conversation is still being followed.


Complex work can continue in the background


GPT Live separates continuous voice interaction from tasks that require deeper reasoning, web search, or more agentic work. When a request becomes more demanding, the voice model can delegate it to a frontier model while maintaining the conversation and returning the result when it is ready.


At launch, GPT Live uses GPT-5.5 for this background work. ChatGPT Voice also offers Instant, Medium, and High intelligence levels when supported by the user's plan, making it possible to choose between faster responses and additional reasoning for more complex questions.


Images and visual responses remain part of the conversation


Voice conversations can continue inside the same chat as text and supported images. A creator can attach a visual reference, type a precise detail when speaking is inconvenient, and continue receiving spoken responses without opening a separate conversation.


ChatGPT can also display supported visual cards while GPT Live is active, allowing certain information to appear on screen instead of being communicated only through audio. Search and memory remain available, although access to individual features can still depend on the account, plan, region, and app version.


Global rollout, models, and initial limitations


GPT-Live-1 and GPT-Live-1 mini began rolling out globally on July 8, 2026, across the ChatGPT apps for iOS and Android and on ChatGPT.com. GPT-Live-1 is becoming the default ChatGPT Voice model for Go, Plus, and Pro users, while Free accounts use GPT-Live-1 mini. OpenAI also plans to bring the models to the API.


Usage is measured over a rolling 24-hour period, with limits depending on the plan, and a single Live conversation can last up to two hours. GPT Live does not initially support video, screen sharing, connected apps, plugins, custom GPTs, Work, or Codex. Eligible mobile subscribers can continue using Advanced Voice Mode when video or screen sharing is needed.


IMPORTANT: GPT Live is being released gradually, so availability may depend on the user's plan, region, workspace, and app version. Voice transcripts may not reproduce every spoken word exactly, especially when speech overlaps or background noise is present.{alertWarning}

Daisuki's Take: What This Means for Designers


GPT Live is most interesting for creative work when speaking becomes part of the thinking process. Designers can talk through an early concept, compare directions, describe a composition, or discuss an attached reference without waiting for a strict sequence of recorded prompts and completed answers.


The ability to delegate research or reasoning while the conversation continues may also reduce interruptions during exploratory work. However, a natural voice does not make every suggestion accurate, and the transcript should not be treated as a precise record of a client meeting, creative brief, or final production decision.


We see GPT Live as a companion for ideation, discussion, and hands-free review rather than a replacement for the visual workspace. Important names, measurements, licensing details, and final instructions should still be confirmed in text before they become part of a finished design.



Sources and Recommended Links