Yesterday's Top Poster

GPT-Live 1 now available on AI Gateway

  • Thread starter Thread starter Kevin Dawkins, Gregor Martynus, Jerilyn Zheng
  • Start date Start date
GPT-Live 1 from OpenAI is now available on AI Gateway.

GPT-Live 1 is a full-duplex voice model and can listen and speak at the same time. Many voice models use turn detection to respond. Full duplex removes that boundary, so a user can pause, interrupt, or add detail while GPT-Live is speaking.

GPT-Live 1 also supports client delegation. Client delegation lets you choose the background model independently from GPT-Live 1. A basic session runs only the voice model. When deeper work is needed, your application can handle delegation-created with any text model on AI Gateway while the conversation continues.

Install AI SDK 7, @ai-sdk/openai 4.0.67 or later, and a WebSocket client:

Simple example without delegation​


This starts openai/gpt-live-1 without calling another model. Wait for session-started before sending audio.

Delegate work and keep talking​


You can call a text model for GPT Live to delegate to, then return the result on the commentary channel for it to speak. This example uses openai/gpt-5.6-sol, but you can substitute any text model available on AI Gateway:


Your application controls delegated work and its permissions, confirmations, and cancellation. Delegated model requests are billed separately through AI Gateway; voice-session usage continues while they run. These snippets omit WebSocket connection, event parsing, transcript assembly, audio streaming, and shutdown. See the GPT-Live guide in the AI Gateway docs for connection setup, audio streaming, and complete examples. For all audio models on AI Gateway, go to the model list.

Read more

Continue reading...
 
Back
Top