OpenAI Unveils GPT-Live to Bridge the Gap Between AI and Human Speech 

Written By: Muskan Saini Published:
GPT live

Communicating with a virtual assistant was traditionally akin to interacting on a digital walkie-talkie: you would talk, pause, and then hear an artificial response. Going against the conventional flow, OpenAI has launched GPT-Live, a duo of voice models from the future generation that enables ChatGPT to simultaneously listen and speak.

The main novelty behind this innovation is the so-called “full-duplex” approach used by engineers. While earlier solutions required complete silence to start answering a user’s question, GPT-Live continuously analyzes incoming audio input several times each second. 

This allows people to interrupt the assistant midway through its speech in order to redirect it without causing any confusion. Moreover, the AI throws in casual human verbal expressions such as “mhmm” and “got it” in order to demonstrate its comprehension of the speech, or remains silent if it detects a mere pause.

Two versions of OpenAI are deployed globally through the iOS, Android, and web interfaces. Free users get an automatic update to the GPT-Live-1 mini version, while those paying for the Go, Plus, and Pro subscriptions get access to the advanced GPT-Live-1 version. 

Notably, the system is capable of conducting web searches live in the middle of the vocal conversation. Upon receiving a challenging question, the voice interface uses background models such as GPT-5.5 to engage in the deep reasoning process.

Although the initial tests showed minor language peculiarities, such as the unusually bookish style of expression and a strong American accent during live tests in languages such as Hindi, this update represents a major breakthrough in design. 

With the voice being transformed into a highly flexible, hands-free user interface that can operate for 40-minute-long discussions, OpenAI is setting the stage for future interactions in which humans will be able to easily control complicated computing processes via natural speech.

About the Author
Muskan Saini

<p><strong>Tech Journalist</strong></p> <p>Muskan is a journalist at Breaking Arc, where she covers stories that influence and inform the readers. Her work is rooted in accuracy, context, and a firm commitment to keeping opinion separate from fact. Her path into journalism began with a Bachelor's and Master's degree in Journalism and Mass Communication, where she built a strong foundation in reporting, media ethics, and the responsibility that comes with the written word. </p> <p>With over a year of experience across News18 and JK24x7, she's developed a sharp eye for detail and an unwavering commitment to truth. Her responsibilities ranged from covering fast-turnaround stories on tech launches and industry updates to learning the discipline of verifying every fact before it reached a reader. </p> <p>Today, at Breaking Arc, Muskan brings that same rigor to covering the technology beat, from business trends to tech innovation; she writes on a wide range of editorial niches with a focus on educating, empowering, and entertaining the audience.</p> <div class="h5">Areas of Coverage</div> <ul class="mb-5"> <li>Technology</li> <li>Artificial Intelligence</li> <li>Consumer Tech</li> <li>Cybersecurity</li> <li>Digital Platforms</li> <li>Software & Apps</li> </ul> <div class="h5">Editorial Standards</div> <p>Muskan follows Breaking Arc’s editorial standards by prioritizing: </p> <ul class="mb-5"> <li>Fact-based and source-backed reporting</li> <li>Independent verification of information</li> <li>Reliance on primary and authoritative sources</li> <li>Timely updates and corrections when required</li> <li>Clear distinction between news, analysis, and opinion</li> </ul> <div class="h5">Language</div> <ul class="mb-5"> <li>English</li> <li>Hindi</li> </ul>

Read more