AI Tool Profile
GPT-Live: ChatGPT Voice capabilities, limits and API status
GPT-Live powers ChatGPT Voice with full-duplex listening and speaking and can delegate search, reasoning and longer tasks to a separate frontier model while the conversation continues. GPT-Live API access has not launched.

Verification & Sources
- Status
- Verified
- Source links
- 4
- Freshness
- Verified July 29, 2026
- Last verified
- July 29, 2026
- Last updated
- July 29, 2026
Key source checks
Suggest a correction
What It Does
GPT-Live powers ChatGPT Voice with full-duplex listening and speaking and can delegate search, reasoning and longer tasks to a separate frontier model while the conversation continues. GPT-Live API access has not launched.
Full Guide
Kingy verdict: GPT-Live is a meaningful redesign of ChatGPT Voice because it separates the live conversation from deeper work. The voice model can keep listening and speaking while another model handles search, reasoning or a longer task. That is more consequential than a nicer synthetic voice: it changes voice from a turn-by-turn input method into a possible control surface for ongoing work. The launch evidence is strong for ChatGPT. It is not evidence of a GPT-Live API release.
What launched—and what did not
OpenAI’s July 8 announcement introduced GPT-Live-1 and GPT-Live-1 mini inside ChatGPT Voice. It says the rollout began globally across iOS, Android and ChatGPT.com, with GPT-Live-1 assigned to Go, Plus and Pro and the mini version assigned to Free. The same announcement says API access is planned “soon.” Kingy therefore treats the product as launched in ChatGPT and unavailable as a GPT-Live API unless a later official API document says otherwise.
Current OpenAI product documentation describes ChatGPT Voice in the desktop app and Remote on iOS, with rollout and workspace controls. That is a newer product surface, not permission to rewrite the July 8 event as an API release. OpenAI’s current pricing documentation is explicit that Desktop Voice is not available through an API key.
How the interaction model differs
Traditional voice pipelines transcribe speech, send text to a language model and synthesize the reply. Turn-based audio models reduce that chain but still wait for a pause before responding. OpenAI describes GPT-Live as full duplex: it processes incoming audio while generating output and repeatedly decides whether to listen, speak, pause, interrupt or invoke a tool. In practical terms, a user should be able to interrupt, think aloud or ask the system to wait without every silence becoming an accidental handoff.
The second change is delegation. GPT-Live manages the live conversation while a frontier model handles deeper work. At the July 8 launch, OpenAI named GPT-5.5 as the background model. Current desktop documentation names GPT-5.6 Terra for task coordination. That difference shows why the voice layer and delegated model should be evaluated separately—and why a profile should not freeze the background model as a permanent GPT-Live specification.
Availability, plans and the API boundary
There is no standalone public GPT-Live price in the reviewed sources. ChatGPT access follows the relevant plan and rollout. OpenAI’s current desktop documentation lists separate voice allowances in rolling five-hour windows and notes that tasks started through Voice also consume the user’s existing Codex budget. Business, Edu and Enterprise workspaces may also be subject to credit-based terms. Limits can change, so the live plan page is the authority.
Developers should not infer an endpoint, model ID or production entitlement from the consumer announcement. GPT-Live is not the same as the existing Realtime API, and “API soon” is not an API launch. Until OpenAI publishes current developer documentation for GPT-Live access, architecture and procurement decisions should use the voice APIs that are actually documented and available.
Evidence and safety boundaries
OpenAI reports preference gains over Advanced Voice Mode and improvements on scientific reasoning, search and a telecom-support evaluation. Those are first-party results. Kingy did not reproduce them, and a polished conversation can conceal errors in delegated work. Teams should inspect the returned sources, preserve task logs where appropriate and keep human approval around sensitive actions.
The system card and launch material discuss voice-specific safety, predefined voices, teen protections and monitoring for emotional reliance. Those controls do not remove ordinary privacy questions. A microphone can capture bystanders, confidential meetings, customer information or screen context. Workspace policy, consent and data-handling review matter before voice is introduced into regulated or shared environments.
A practical evaluation plan
- Run the same task in quiet, background speech and intermittent noise; record false interruptions, missed corrections and recovery quality.
- Test natural pauses, mid-sentence interruption and an explicit request to stay silent. Score whether the system follows the conversational instruction.
- Give one task that requires delegation. Verify the result, sources, latency and whether the live conversation accurately reports progress.
- Repeat in the languages and accents that matter to the actual users; do not generalize from an English demo.
- Check plan limits, workspace controls, microphone permissions, screen context and any retained transcript before deployment.
- For developer use cases, confirm a documented API product separately. Do not treat ChatGPT access as an API entitlement.
Primary sources
Tool Links
Launch History
GPT-Live launches in ChatGPT Voice; API access remains pending
OpenAI began rolling GPT-Live-1 and GPT-Live-1 mini into ChatGPT Voice for users globally. The voice model uses full-duplex audio for simultaneous listening and speaking, and it can delegate…
GPT-Live’s important change is architectural and experiential: continuous interaction is separated from deeper task execution. That could make voice useful for steering work rather…