OpenAI Officially Rolls Out GPT-Live-1 API: Enabling Interruption Management, Tool Invocation, and Telephony Agents
1 day ago / Read about 0 minute
Author:小编   

OpenAI has unveiled a cutting-edge model, GPT-Live-1, along with the release of its API. This innovative model seamlessly integrates speech comprehension and generation, departing from the conventional multi-step connection framework. This integration minimizes system latency and enables full-duplex interactions, allowing for precise handling of intricate conversational situations. With its versatile functional expansion and scenario adaptability, developers have the flexibility to fine-tune attributes like tone, integrate tools, or craft voice-driven workflows. This makes it an ideal fit for applications such as telephony services and restaurant reservations.

In real-world deployments, GPT-Live-1 significantly cuts down on erroneous interruptions, simplifies the system architecture, and showcases outstanding performance in evaluations. The GPT-Live-1 API is now accessible, with transparent pricing for the front-end voice component. Additionally, OpenAI has introduced a selection of new voices and has plans to broaden its support for an even wider array of voices and languages moving forward.