Build more natural voice experiences with GPT‑Live‑1 in the API
GPT-Live-1 adds full-duplex voice, improved instruction following, custom voices, and telephony support for builders shipping conversational audio products.
**GPT-Live-1** brings **full-duplex voice conversations** to the API, alongside stronger instruction following, custom voices, and telephony support.
Voice-agent builders should reconsider architectures that separate turn detection, response generation, and playback when a live, bidirectional API can handle the conversational loop more directly.
**GPT-Live-1** brings **full-duplex voice conversations** to the API, alongside stronger instruction following, custom voices, and telephony support. Voice-agent builders should reconsider architectures that separate turn detection, response generation, and playback when a live, bidirectional API can handle the conversational loop more directly. The material provides no latency, pricing, language coverage, safety controls, or migration details, so production tradeoffs cannot yet be assessed from it.
This shifts the voice-agent design choice from assembling transcription, turn handling, generation, and playback toward a managed full-duplex conversational loop. It confirms that bidirectional speech APIs are becoming a distinct model surface, but missing latency, interruption, safety, language, and cost evidence prevents deciding whether the integrated loop should replace modular audio pipelines.