GPT-Live-1 Comes to the API: What It Means for Voice Agents
The model can listen and speak simultaneously, delegate deeper work, and support telephony—but the advertised voice-layer rate is not the complete cost of an agent.

Bottom line
GPT-Live-1 is now available to developers at $0.05 per minute for the front-end voice layer. Here is what full-duplex changes and what teams still need to test.
Editorial accountability
Who checked this guide
- Evaluation type
- Research-based verification
- Last materially checked
- Evidence
- 4 listed sources
Hands-on testing is identified explicitly. Research-based coverage uses cited product documentation and other named sources; it does not imply every paid plan was used. Read the full methodology.
Editorial basis
What this guidance is based on
- Editorial basis
- Source-led analysis
- Primary references
- 4
- Products covered
- 1
- Last checked
- 2026-09-11
Important limits
- • Features, availability, and pricing can change after publication; confirm consequential details with the provider.
In this guide
*This research-based analysis covers OpenAI's September 10, 2026 API release. Performance figures and early-customer results are OpenAI's reported evaluations unless otherwise stated. Pricing and capabilities can change.*
The short answer
GPT-Live-1 is now available through the OpenAI API as a full-duplex voice layer priced at $0.05 per minute. It can listen while speaking, handle interruptions and brief acknowledgments, produce transcripts and response text, support telephony, and delegate reasoning or tool calls to a separate backend model such as GPT-6 Astra.
The release targets a persistent weakness in voice agents: chained speech-to-text, language-model, and text-to-speech systems often lose timing and conversational context at each handoff. A single front-end model may make turn-taking feel more natural. It does not eliminate backend-model charges, telephony, tools, infrastructure, monitoring, or human escalation.
What is full-duplex voice?
Full duplex means the system can process incoming speech while producing outgoing audio instead of forcing rigid alternating turns. That enables a caller to interrupt, correct a detail, laugh, hesitate, or give a short acknowledgment without the entire interaction resetting.
OpenAI says GPT-Live-1 reasons across incoming and outgoing audio together, manages silence and background noise, and retains quality across longer sessions. It also exposes transcripts, keyword biasing, and turn detection for applications that still need explicit boundaries.
How much does GPT-Live-1 cost?
OpenAI lists $0.05 per minute for the front-end voice layer. A ten-minute conversation therefore starts at $0.50 for that layer. The full cost can also include the delegated text or reasoning model, agent harness, tool calls, phone carrier, storage, observability, and review.
Model cost per resolved conversation, not merely per minute. Track containment rate, transfers, repeated questions, tool errors, user abandonment, review labor, and any credits or remediation caused by a bad interaction.
What can developers build?
The launch describes browser connections through WebRTC, server integrations through WebSockets, and phone workflows through telephony. Likely uses include support, reservations, tutoring, intake, internal help desks, and voice access to a backend agent. OpenAI also offers a sales-led enterprise layer called Presence.
Delegation creates a useful architecture: GPT-Live-1 manages conversational timing while a backend model or tool handles deeper work. It also creates a handoff that must be tested for delay, duplicated actions, stale context, and what the voice layer says while the backend is still working.
What should teams test before launch?
Evaluate real accents, dialects, speech differences, code-switching, names, account numbers, noisy rooms, weak connections, long silences, emotional callers, interruptions, and simultaneous speech. Test hang-ups and reconnects, tool timeouts, transfers, identity verification, consent, recording notices, emergency language, and explicit confirmation before consequential actions.
Users should know they are speaking with AI. Custom voices require appropriate rights and consent, and organizations need policies for recordings, transcripts, retention, sensitive information, and access. A natural voice raises the risk of misplaced trust; disclosure and escalation should become clearer as realism improves.
The verdict
GPT-Live-1 makes a credible architectural change by separating natural full-duplex conversation from the backend intelligence and tools a workflow needs. The published $0.05-per-minute rate makes early cost modeling straightforward, but it is only the front layer.
Pilot one narrow workflow where interruptions matter, compare it with the current system, and require safe escalation. The goal is not the most human-sounding demo; it is a correctly resolved conversation with informed users, controlled actions, and a complete evidence trail.
Sources and verification
Product details and claims were checked against the following primary sources.
Frequently asked questions
What is GPT-Live-1?
GPT-Live-1 is OpenAI's API voice model for full-duplex conversations. It can listen and speak simultaneously, handle interruptions, produce transcripts, and delegate deeper work to backend models and tools.
How much does GPT-Live-1 cost?
OpenAI lists $0.05 per minute for the front-end voice layer. Backend models, tools, agent infrastructure, telephony, storage, and monitoring can add to the total.
Does GPT-Live-1 support phone calls?
Yes. OpenAI says it supports telephony, along with WebRTC for browser applications and WebSockets for server integrations.
Is GPT-Live-1 safe for customer service?
It can support customer service, but teams should test identity, consent, interruptions, tool errors, consequential-action confirmation, disclosure, retention, and reliable transfer to a person before production use.
Found this useful?
Get the next one in your inbox.
One five-minute briefing a week: a meaningful change, a practical workflow, and a clearer tool decision—already filtered for lean teams.
Free · one email a week · unsubscribe any time
Tools mentioned in this article
ChatGPT
The general-purpose AI assistant that started it all
OpenAI's flagship conversational AI model, powering everything from casual chat to complex reasoning, coding, and creative work.
Read next
Recommended for you
NotebookLM Review 2026: Is Google's Research Assistant Worth Using?
A research-based review of NotebookLM's source-grounded chat, citations, generated study formats, privacy boundaries, and fit for serious knowledge work.
NotebookLM is one of the best AI research tools for interrogating a defined source library, but its citations improve verification rather than eliminating it—and it is not a replacement for finding or judging the evidence.
Read guide
ChatGPT Images 2.5: What Changed and Who Should Use It
Best AI Tool for Nonprofits in 2026: Canva Is Our Top Pick
Grok Voice Think Fast 2.0 Is Now the Default: What xAI's Voice Upgrade Means for Speech AI