OpenAI has released GPT‑Live‑1 directly in the API so developers can build voice-enabled apps and business workflows with the same natural spoken model that powers ChatGPT. The model can listen and speak at the same time (full-duplex), and it can delegate deeper reasoning and tool actions to backend models such as GPT‑6 Astra.

Why it matters for voice agent builders

Traditional voice agents chain speech-to-text, a reasoning model, and text-to-speech. Every handoff adds latency and breaks the natural rhythm of a conversation. GPT‑Live‑1 handles listening and speaking in a single model, responding to interruptions and acknowledgements in real time while work happens in the background. Developers can also shape an agent’s tone, pace, and conversational style through the system prompt.

Key strengths announced

Real-world results

New voices and pricing

OpenAI expanded the voice catalog across accents, dialects, and languages (Quartz, Ripple, Vesper, Willow, Stone, Gleam, Meridian, Bossa, Tempo, Beacon, Delta, Cinder, and more), with additional options planned. GPT‑Live‑1 is available today in the API at $0.05 per minute for the front-end voice layer. Enterprises can also build on it through OpenAI Presence for trusted, action-taking agents.

Suijin Greenverse view

Voice agents are moving from stop-start chatbots toward natural phone-grade conversations. The interesting architecture shift here is not the voice itself — it is the delegation boundary: the conversation engine stays fast and human, while real reasoning and actions happen behind it. Teams building customer-facing voice workflows should evaluate interruption handling first; it is where the difference is most visible.

Source: OpenAI, September 10, 2026

Leave a Reply

Your email address will not be published. Required fields are marked *

Trang chủ

Nhân vật

Về Greenverse