01 What happened
On September 10, 2026, OpenAI made GPT-Live-1 available via API for full-duplex voice conversations that listen and speak simultaneously.
02 Key details
- The model manages interruptions and pauses natively, delegating deeper reasoning and tool usage to a developer-selected backend model.
- Developers can define tone, pace, and style using system instructions for voice agents, including automated phone workflows.
- The voice layer costs $0.05 per minute, while backend model and tool usage incur separate charges.
- OpenAI reports that early Speak evaluations demonstrated nearly 80% fewer interruptions during thinking pauses compared to previous turn-based systems.
03 Why it matters
This API allows developers to engineer responsive voice agents and phone workflows that maintain a natural conversational flow by handling interruptions and pauses in real-time.
04 Who it matters to
Software developers and voice system engineers.
Original sourceOpenAI