GuideUpdated 2026-09-11

GPT-Live-1 Comes to the API: What It Means for Voice Agents

The model can listen and speak simultaneously, delegate deeper work, and support telephony—but the advertised voice-layer rate is not the complete cost of an agent.

By DiscoverAI Editorial TeamReviewed by DiscoverAI Editorial Review3 min readVideo, Audio & CreativeHow we evaluate
Paper-cut person and AI voice interface exchanging overlapping audio waves with protected tools in the background
Original DiscoverAI editorial illustration. Editorial illustration: natural turn-taking is the front end; safe tools, consent, escalation, and accountable outcomes complete the system.

Bottom line

GPT-Live-1 is now available to developers at $0.05 per minute for the front-end voice layer. Here is what full-duplex changes and what teams still need to test.

Editorial accountability

Who checked this guide

Meet the editorial team →
Evaluation type
Research-based verification
Last materially checked
Evidence
4 listed sources

Hands-on testing is identified explicitly. Research-based coverage uses cited product documentation and other named sources; it does not imply every paid plan was used. Read the full methodology.

Editorial basis

What this guidance is based on

Editorial basis
Source-led analysis
Primary references
4
Products covered
1
Last checked
2026-09-11

Important limits

  • Features, availability, and pricing can change after publication; confirm consequential details with the provider.
In this guide
  1. The short answer
  2. What is full-duplex voice?
  3. How much does GPT-Live-1 cost?
  4. What can developers build?
  5. What should teams test before launch?
  6. The verdict

*This research-based analysis covers OpenAI's September 10, 2026 API release. Performance figures and early-customer results are OpenAI's reported evaluations unless otherwise stated. Pricing and capabilities can change.*

The short answer

GPT-Live-1 is now available through the OpenAI API as a full-duplex voice layer priced at $0.05 per minute. It can listen while speaking, handle interruptions and brief acknowledgments, produce transcripts and response text, support telephony, and delegate reasoning or tool calls to a separate backend model such as GPT-6 Astra.

The release targets a persistent weakness in voice agents: chained speech-to-text, language-model, and text-to-speech systems often lose timing and conversational context at each handoff. A single front-end model may make turn-taking feel more natural. It does not eliminate backend-model charges, telephony, tools, infrastructure, monitoring, or human escalation.

What is full-duplex voice?

Full duplex means the system can process incoming speech while producing outgoing audio instead of forcing rigid alternating turns. That enables a caller to interrupt, correct a detail, laugh, hesitate, or give a short acknowledgment without the entire interaction resetting.

OpenAI says GPT-Live-1 reasons across incoming and outgoing audio together, manages silence and background noise, and retains quality across longer sessions. It also exposes transcripts, keyword biasing, and turn detection for applications that still need explicit boundaries.

How much does GPT-Live-1 cost?

OpenAI lists $0.05 per minute for the front-end voice layer. A ten-minute conversation therefore starts at $0.50 for that layer. The full cost can also include the delegated text or reasoning model, agent harness, tool calls, phone carrier, storage, observability, and review.

Model cost per resolved conversation, not merely per minute. Track containment rate, transfers, repeated questions, tool errors, user abandonment, review labor, and any credits or remediation caused by a bad interaction.

What can developers build?

The launch describes browser connections through WebRTC, server integrations through WebSockets, and phone workflows through telephony. Likely uses include support, reservations, tutoring, intake, internal help desks, and voice access to a backend agent. OpenAI also offers a sales-led enterprise layer called Presence.

Delegation creates a useful architecture: GPT-Live-1 manages conversational timing while a backend model or tool handles deeper work. It also creates a handoff that must be tested for delay, duplicated actions, stale context, and what the voice layer says while the backend is still working.

What should teams test before launch?

Evaluate real accents, dialects, speech differences, code-switching, names, account numbers, noisy rooms, weak connections, long silences, emotional callers, interruptions, and simultaneous speech. Test hang-ups and reconnects, tool timeouts, transfers, identity verification, consent, recording notices, emergency language, and explicit confirmation before consequential actions.

Users should know they are speaking with AI. Custom voices require appropriate rights and consent, and organizations need policies for recordings, transcripts, retention, sensitive information, and access. A natural voice raises the risk of misplaced trust; disclosure and escalation should become clearer as realism improves.

The verdict

GPT-Live-1 makes a credible architectural change by separating natural full-duplex conversation from the backend intelligence and tools a workflow needs. The published $0.05-per-minute rate makes early cost modeling straightforward, but it is only the front layer.

Pilot one narrow workflow where interruptions matter, compare it with the current system, and require safe escalation. The goal is not the most human-sounding demo; it is a correctly resolved conversation with informed users, controlled actions, and a complete evidence trail.

Sources and verification

Product details and claims were checked against the following primary sources.

Frequently asked questions

What is GPT-Live-1?

GPT-Live-1 is OpenAI's API voice model for full-duplex conversations. It can listen and speak simultaneously, handle interruptions, produce transcripts, and delegate deeper work to backend models and tools.

How much does GPT-Live-1 cost?

OpenAI lists $0.05 per minute for the front-end voice layer. Backend models, tools, agent infrastructure, telephony, storage, and monitoring can add to the total.

Does GPT-Live-1 support phone calls?

Yes. OpenAI says it supports telephony, along with WebRTC for browser applications and WebSockets for server integrations.

Is GPT-Live-1 safe for customer service?

It can support customer service, but teams should test identity, consent, interruptions, tool errors, consequential-action confirmation, disclosure, retention, and reliable transfer to a person before production use.

Found this useful?

Get the next one in your inbox.

One five-minute briefing a week: a meaningful change, a practical workflow, and a clearer tool decision—already filtered for lean teams.

Free · one email a week · unsubscribe any time

Tools mentioned in this article

ChatGPT

The general-purpose AI assistant that started it all

4.6

OpenAI's flagship conversational AI model, powering everything from casual chat to complex reasoning, coding, and creative work.

FreemiumChatbotsWriting

Read next

More on Video, Audio & Creative