Good news!
"OpenAI releases full duplex voice model to the API
OpenAI released GPT-Live-1, a voice model that listens and speaks simultaneously rather than chaining separate speech-to-text, reasoning, and text-to-speech models.
It handles interruptions and background noise within a single model, delegating deeper reasoning and tool calls to a backend such as GPT-6 Astra, and supports telephony deployments for phone-based customer support.
OpenAI reports a 30-percentage-point improvement on Full Duplex Bench over GPT-Realtime-2.1 and first place on Tau3 when paired with GPT-6 Astra;
Speak’s early evaluation showed an 80 percent cut in interruptions versus turn-based systems. Pricing is $0.05 per minute for the voice layer, with backend model and tool costs billed separately; it natively supports ASR transcripts, keyword biasing, and turn detection. For voice app developers, it removes the latency and coordination overhead of stitching together separate STT, LLM, and TTS components." (Data Points)
No comments:
Post a Comment