OpenAI made GPT Live 1 generally available in its API on September 10, 2026. The full duplex voice model can listen and speak concurrently, while a separate backend handles reasoning and tools; the voice layer costs $...
Published byEdited with GPT-5.6 TerraImages generated with GPT Image 2
Research answer

Create a landscape editorial hero image for this Studio Global article: What did OpenAI announce on September 11, 2026, about launching the GPT-Live-1 full-duplex voice model in its API, including how it differs. Article summary: OpenAI’s API announcement was dated September 10, 2026—not September 11. It introduced GPT‑Live‑1 as a generally available, full‑duplex voice layer: a model intended to make voice agents more natural by listening and spe. Topic tags: general, general web, documentation. Style: premium digital editorial illustration, source-backed research mood, clean composition, high detail, modern web publication hero. Use reference image context only for broad subject, composition, and topical grounding; do not copy the exact image. Avoid: logos, brand marks, copyrighted characters, real person likenesses, fake screenshots, UI text, readable text, watermarks, charts with fake
OpenAI announced the general availability of GPT-Live-1 in its API on September 10, 2026, not September 11. The release gives developers a voice layer for building agents that can listen and speak at the same time, instead of treating a conversation as a sequence of rigid, completed turns. 1
16
Many voice-agent systems are built as a chain: speech-to-text converts the caller’s audio into text, a language model produces a response, and text-to-speech turns that response back into audio. GPT-Live-1 is positioned as a unified voice layer for the real-time conversational part of that interaction. 1
The practical distinction is turn-taking. OpenAI says GPT-Live-1 can listen and speak simultaneously, follow pauses and shifts in pace, and decide whether it should answer or continue listening. That design is intended to make interruptions and brief acknowledgements—often called backchannels—feel less like errors and more like normal conversation. 1
9
15
This does not mean the voice model must perform every task itself. A GPT-Live-1 conversation can continue while a backend model or agent performs deeper reasoning or uses tools. OpenAI supports Responses delegation with an OpenAI model and client delegation for a developer’s own backend, enabling teams to choose the reasoning system behind the voice experience. 1
16
GPT-Live-1 acts as the front-end conversational voice layer. The backend can be responsible for tasks such as more involved reasoning, retrieval, workflow execution, or tool calls. In OpenAI’s own product notes, GPT-6 Astra is identified as an option for harder search and reasoning tasks; API developers can also connect their own backend through client delegation. 4
16
For builders, that separation creates a trade-off:
OpenAI also highlighted stronger instruction following, custom voices, and telephony support for the API release. 13
GPT-Live-1 voice sessions cost $0.05 per minute. Billing is calculated per second, with no rounding up to the next full minute. 16
18
That is the price for the voice session—not necessarily the complete cost of a voice agent. Backend model usage and tool usage are billed separately, so a production budget needs to account for both the duration of the conversation and the work delegated behind it. 16
18
OpenAI reported that GPT-Live-1 improved by 30 percentage points on its Full Duplex Bench compared with GPT-Realtime-2.1, and that GPT-Live-1 paired with GPT-6 Astra at medium reasoning effort ranked first on Tau3. 1
However, the supplied source material does not independently substantiate the specific figures in the original question for turn-taking latency or the exact Tau3 pass-rate comparison. Those precise numbers should not be treated as verified here.
OpenAI described several early production deployments as evidence for the value of more natural turn-taking:
These are company-reported outcomes rather than independent evaluations, so they are best read as deployment examples—not universal performance guarantees.
The available evidence does not substantiate claims about enterprise access through a managed product called “Presence,” nor does it establish a specific expansion in accents, dialects, or languages for this API launch. Those details should be verified directly in current OpenAI documentation before being used in procurement or implementation decisions.
GPT-Live-1’s main change is architectural: it brings a full-duplex conversational voice layer to the API, allowing an agent to keep listening and speaking naturally while a separate system performs slower reasoning or calls tools. At $0.05 per minute for the voice layer, billed by the second, it also makes the cost of live voice time straightforward—but the complete cost still depends on the selected backend model and tools. 1
16
18
Studio Global AI
This page includes a source-backed answer you can continue inside Studio Global.
OpenAI made GPT Live 1 generally available in its API on September 10, 2026. The full duplex voice model can listen and speak concurrently, while a separate backend handles reasoning and tools; the voice layer costs $...
OpenAI made GPT Live 1 generally available in its API on September 10, 2026. The full duplex voice model can listen and speak concurrently, while a separate backend handles reasoning and tools; the voice layer costs $... Rather than forcing a user to finish a turn before responding, GPT Live 1 is designed to handle pauses, interruptions, acknowledgements, and changes in conversational direction in real time.
The supplied evidence supports OpenAI’s broad claims about full duplex benchmarks and early production use, but not every precise latency, pass rate, enterprise access, or language coverage claim.