OpenAI's conversational voice model released on 8 July 2026, built on a full-duplex architecture that listens and speaks simultaneously. It replaces Advanced Voice Mode in ChatGPT for the paid Go, Plus and Pro tiers, and delegates queries requiring search or reasoning to GPT-5.5 running in the background.
Release date
8 July 2026
Access:HostedDeployment:☁ Cloud
Overview
Applications
Access & deployment
Hosted
Cloud
Weights: Closed
Key parameters
✓ Tools
📥 Input: text, audio
Technical specification
License
Proprietary
Features:✓ Tool use
Modalities
⬇ Input
textaudio
⬆ Output
textaudio
Capabilities and applications
Native model capabilities
Voice Conversation
Ability to conduct multi-turn real-time voice conversations with context retention and natural speech pacing.
Category: speech
Speech to text
Category: speech
Text to speech
Category: speech
Real-time inference
The model's ability to generate responses with very low latency (>1000 tokens/sec) on specialized inference hardware (e.g. Cerebras WSE), enabling interactive, turn-by-turn collaboration with a human.
Category: coding
Streaming Speech-to-Text
Real-time conversion of speech to text with immediate output as the speaker is talking.
Category: speech
Live Translation
Real-time speech translation between multiple languages without interrupting the audio stream.
Category: speech
Multilingual
Competence in many natural languages (from a few to over a hundred): understanding, generation, translation, and code-switching within a single conversation. Frontier models support a wide range of languages with comparable quality.
Category: language
Application domains
Benchmark results
3 benchmarks
GPQA
accuracy · GPT-Live-1 substantially outperforms Advanced Voice Mode; exact score not publicly disclosed by OpenAI.
📄 OpenAI blog — Introducing GPT-Live
Expert-level scientific reasoning benchmark. OpenAI published only a comparison chart, without the raw score.
BrowseComp
accuracy · Better than Advanced Voice Mode; exact number not published.
📄 OpenAI blog — Introducing GPT-Live
Tests agentic web search and the ability to find hard-to-locate information.
τ³-Voice Telecom (internal variant)
task success · GPT-Live-1 outperforms Advanced Voice Mode; numeric score not disclosed.
📄 OpenAI blog — Introducing GPT-Live
Realistic, multi-turn voice telecom support scenarios. Customized user model variant.
Pricing
Technical architecture
Core Architecture
Model Form
Training Techniques
Deployment and security
🔒 Security / Enterprise
✓ Verified enterprise information
Updated: 8 Jul 2026↗ Security documentation
