OpenAI put its full-duplex voice model, GPT-Live-1, into the hands of developers on September 10, charging five cents a minute for the voice layer. The pitch is simple: a model that can listen and speak at the same time, the way people actually hold a conversation.
Turn-based voice assistants wait for you to finish, process, then reply. GPT-Live-1 does not wait its turn. It can talk while you talk, notice when you are pausing to think rather than finishing, break in when that is appropriate, and call tools mid-sentence, according to OpenAI's announcement. The company frames the voice model as a front end that hands deeper reasoning to a separate text model behind it, such as GPT-6 Astra or a third-party model of the developer's choosing.
The numbers behind the feel
OpenAI says GPT-Live-1 improves by 30 percentage points on its Full Duplex Bench over the earlier GPT-Realtime-2.1, with most of the gain in turn-taking and latency. The language-learning app Speak, an early tester, found that the model cut false interruptions, the moments where an assistant talks over a learner who is just thinking, by nearly 80 percent compared with a traditional turn-based setup, as reported by Unite.AI.
Developers can steer the voice through system prompts, adjusting tone, speaking speed, and conversational style, then delegate the hard thinking to a backend model. That split keeps the responsive voice layer cheap while letting the expensive reasoning run only when it is needed.
Why it matters
Getting the rhythm of speech right is the part that has held voice AI back. A system that interrupts you, or freezes while it thinks, breaks the illusion of talking to something that understands you. Closing that gap is what turns a voice demo into something people will use for customer support or language tutoring.
The release also extends OpenAI's strategy of selling the GPT-6 Astra family as the brain and GPT-Live-1 as the mouth and ears, a division of labour that lets each piece improve on its own schedule. Rivals building voice agents now have a concrete latency and interruption benchmark to beat.
Sources
- i. openai.com
- ii. www.unite.ai
- iii. www.testingcatalog.com
Commentarii · 0