This page summarizes the release announcement.
--
OpenAI's GPT-Live is a next-generation voice model (powered by ChatGPT Voice) designed for natural, real-time voice conversations. Its key features and functions are as follows:
1. Continuous Interaction
Full-Duplex Architecture: Unlike previous models that waited for the user's speech to finish completely, GPT-Live listens in real time and can speak simultaneously.
Natural Conversation Flow: During a conversation, GPT-Live naturally waits if the user interrupts or pauses to think. It also provides active listening responses like "um" and "aha," creating a feeling similar to talking to a real person.
2. Delegation for Deeper Work
Hybrid Structure: GPT-Live, responsible for real-time conversations, and Frontier models, handling complex reasoning and web searches, operate separately.
Background Task Processing: When complex questions or web searches are needed, GPT-Live delegates the task to the latest Frontier model (GPT-5.5 as of release) in the background. This allows for uninterrupted conversation with the user while the model retrieves information and thinks.
3. Three Reasoning Modes Offered
Users can choose the reasoning stage to operate in the background based on their needs.
4. Visual Answers Provided
During voice conversations, information requiring visual confirmation, such as weather, stocks, sports scores, or maps, is presented on the screen in the form of Rich Visual Cards.
5. Enhanced Listening and Noise Cancellation
Improved to focus better on the user's voice even with background noise like vehicle sounds or third-party conversations.
Enhanced ability to wait patiently when the user pauses to think, preventing premature conversation termination.
6. Strengthened Voice Safety
Includes a real-time guardrail that detects and redirects potentially harmful outputs during real-time conversations, either steering towards safe responses or ending the conversation.
Limited to using nine unique voices to prevent voice impersonation.
--
Release and Availability Information: GPT-Live is globally available through iOS, Android, and ChatGPT.com. Paid users (Plus and Pro) have access to GPT-Live-1 as the default voice model, while free users use GPT-Live-1 mini. (Note: Screen sharing and video functionality are not supported at launch but will be added in future updates.)
▶ Original Source: https://openai.com/index/introducing-gpt-live/