GPT-Live 1
OpenAI's full-duplex voice frontend for natural conversation that delegates deeper reasoning and actions to a separate backend agent.
OpenAI's full-duplex voice frontend for natural conversation that delegates deeper reasoning and actions to a separate backend agent.
Plain-English overview
GPT-Live 1 listens and speaks at the same time, so it can handle pauses, acknowledgements, interruptions, and background noise without forcing every exchange into rigid turns. It also provides transcripts and response text alongside the live audio.
The architecture separates conversation from deeper work. GPT-Live 1 manages the voice experience while a Responses model or an application-operated agent handles reasoning and approved tools. The published voice price therefore excludes backend model usage, tools, telephony, and application infrastructure.
Category comparison
These are provider-published specifications, not Cody benchmark scores. Follow the linked sources for current limits and endpoint-specific exceptions.
API and provider access
Availability
Available through OpenAI's Live API for browser, server, and telephony voice applications.
Direct access follows OpenAI's live supported-country list; the Free API tier is not supported for this model.
Check live availabilityData and training
OpenAI says API content is not used for training by default. Default abuse-monitoring logs may be retained up to 30 days, with additional controls available to qualifying organizations.
This is a concise reading of the cited provider material, not legal advice. A third-party gateway can have different storage, routing, training, and residency terms from the model maker's direct API.
Read the provider policyFrequently asked questions
OpenAI's full-duplex voice frontend for natural conversation that delegates deeper reasoning and actions to a separate backend agent. GPT-Live 1 listens and speaks at the same time, so it can handle pauses, acknowledgements, interruptions, and background noise without forcing every exchange into rigid turns. It also provides transcripts and response text alongside the live audio.
GPT-Live 1 was released on September 10, 2026 according to the cited provider materials.
Available through OpenAI's Live API for browser, server, and telephony voice applications. The access routes listed in this guide are OpenAI.
$0.05 / voice minute. Voice sessions are billed per second without whole-minute rounding. Backend model and tool usage is billed separately, as are telephony and application infrastructure.
Available through OpenAI's Live API for browser, server, and telephony voice applications. Direct access follows OpenAI's live supported-country list; the Free API tier is not supported for this model.
OpenAI says API content is not used for training by default. Default abuse-monitoring logs may be retained up to 30 days, with additional controls available to qualifying organizations. The policy belongs to the provider route and account terms, so verify it again before production use.
Related comparisons
OpenAI's realtime speech-to-speech reasoning model for tool-using voice agents that also need text and image context.
Open full comparisonGoogle's preview audio-to-audio model for low-latency dialogue with multimodal awareness, thinking, search grounding, and function calling.
Open full comparisonElevenLabs' expressive realtime text-to-speech model for natural dialogue, emotional delivery, audio tags, and more than 70 languages.
Open full comparisonInworld's newest expressive realtime text-to-speech model, designed to remember conversational delivery and speak in more than 100 languages.
Open full comparisonSpaceXAI's current realtime speech-to-speech model for sub-second conversational agents with tool access.
Open full comparison