
What happened
OpenAI is making GPT-Live-1, a full-duplex speech model that listens and talks at once, available to developers as an API at $0.05 per minute. Yelp already uses it for phone reservations.
Why it matters
On OpenAI's benchmarks, full-duplex interactivity rose to 80.1 percent from 45.4 percent for GPT-Realtime-2.1, while turn-taking latency fell to 0.8 seconds from 1.4 seconds.
What to watch
Whether the $0.05 per minute price holds back adoption for high-volume calling, and whether external testing matches OpenAI's own benchmark gains. Watch the tool-calling accuracy improvement, 87 percent from 60 percent.
WHO IT HITSDevelopers building voice agents and contact-center teams are the clearest beneficiaries, since they can now pair full-duplex speech with different backend models to trade off reasoning depth, speed, and cost. Businesses weighing phone-based customer service will need to justify the $0.05 per minute cost against the reported call-handling gains.
Ask the AI about this article →
Summaries like this, in your inbox every morning.
GPT-Live-1 is not a brand-new model launch so much as a packaging decision: the speech model is already running inside ChatGPT, and OpenAI is now exposing it through an API so developers can build their own voice applications. The company is letting developers pair the speech layer with different backend models depending on the task, matching reasoning depth, speed, and cost to each use case.
The benchmark story is the pitch. Against GPT-Realtime-2.1, OpenAI reports full-duplex interactivity jumping to 80.1 percent from 45.4 percent, turn-taking latency dropping to 0.8 seconds from 1.4 seconds, and tool-calling accuracy rising to 87 percent from 60 percent. In a banking voice support benchmark, GPT-Live-1 reaches a 32 percent pass rate, up from 12.4 percent. These are OpenAI's own numbers, so the test is whether outside developers see the same gains.
The early customer example is Yelp, which is using the model for phone-based reservations and, according to CTO Alex Levy, reports better call handling. For companies weighing voice agents, the outcome likely hinges on whether that kind of operational improvement justifies $0.05 per minute at call volumes, and on whether the twelve new voices and built-in ASR transcripts and response text make integration straightforward enough to adopt.
For example, today's edition would include:
AI-summarized, only the topics you pick — one digest a day via Email, Slack, or Discord.
Free · 30 seconds with Google · unsubscribe anytimeWhat is AIToday? →
Ask AI anything about this article. Q&As are published on this page for other readers too.
Dynatrace acquired Arize AI, adding AI observability, evaluation and agent monitoring to its application obser…
A Daily Dose of Data Science test kept LoRA adapters separate from a shared 7B base model, cutting 100 fine-tu…

A report by Spencer Kitts, Thomas Larsen and Sydney Von Arx says an OpenAI agent swarm very likely ran an atta…

Simon Willison wrote that many people, himself included, have gone through an existential crisis when a coding…

Stephen Aarons, a New Mexico defense lawyer of over 40 years, was held in direct contempt and fined $5,000 for…

Perplexity cofounder and Chief Strategy Officer Johnny Ho said GPT‑6 Astra can craft communications, edit real…
