OpenAI launches GPT-Live-1 full-duplex voice model in the API
OpenAI launched (Sept 10) GPT-Live-1 in the API—the full-duplex voice model first introduced in ChatGPT Voice—so apps can listen and speak at the same time, handle interruptions and brief acknowledgments, and keep talking while a separately chosen backend model/agent (including GPT-6 Astra, Codex, or ChatGPT Work) does deeper reasoning and tool work. Developer controls include tone/pace via prompting, expanded voices, built-in transcripts, keyword biasing, and turn detection; connections cover WebRTC (browsers), WebSockets (servers), and telephony. OpenAI cites ~0.798s response latency vs ~1.41s for GPT-Realtime-2.1, 97.3% on Artificial Analysis Conversational Dynamics, and strong Full Duplex Bench interactivity/tool-calling scores when paired with a backend. Pricing: $0.05/minute for the voice layer, billed per second; backend model and tool usage billed separately. Distinct from the earlier ChatGPT Voice GPT-Live consumer launch, from GPT-Live-Transcribe / GPT-Transcribe STT models, and from GPT-Live SynthID audio provenance.






