OpenAI has launched a version of its latest real-time full-duplex voice model as an API, allowing software developers to implement more responsive voice-enabled apps and conversational workflows. GPT-Live-1 debuted in July, offering a way to interact through spoken prompts and responses instead of through typed text. Initially available through the ChatGPT interface, the voice model can now be reached via API calls. As a full-duplex model, GPT-Live-1 can listen and generate speech at the same time, which aids fluid communication. People often talk over one another and may find it frustrating to take turns speaking and listening as if using half-duplex handheld radios. OpenAI says it has focused on allowing developers implementing...
Läs hela artikeln hos källan.
Kommentarer (0)
Inga kommentarer ännu. Bli först med att kommentera!