415.tech
AI & tech, from the frontlines of Silicon Valley
OpenAI ships full-duplex GPT-Live-1 voice API at $0.05 per minute

OpenAI ships full-duplex GPT-Live-1 voice API at $0.05 per minute

OpenAI made GPT-Live-1 generally available in its API at $0.05 per minute — a full-duplex model that listens and speaks at once, handles interruptions, and routes deeper reasoning to a backend model the developer picks. On OpenAI's benchmarks turn-taking latency drops to 0.8 seconds from 1.4 and tool-calling accuracy rises to 87% from 60%, so a developer can build phone agents that field interruptions the way people do — the pattern Yelp runs for reservations.

Source: the-decoder.com

Post on XEmail

The speech model can listen and talk at the same time, a feature known as "full-duplex," and is already running inside ChatGPT.

The Decoder

Why this matters

  • → Full-duplex voice reduces agent latency 43% and tool-calling accuracy jumps to 87%, enabling natural phone int
  • → Developers can now pair voice with custom backend models to optimize reasoning depth vs. cost per use case.
  • → Production deployments (Yelp) already report measurable improvements in call handling quality.
Voice goes two-way
Also in this edition