Gemini 3.8 Live models now available on AI Gateway
Google's Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking models are now available through Vercel's AI Gateway, supporting real-time spoken interactions via the realtime API with features like multi-language support and parallel reasoning.
and from Google are now available on AI Gateway.Gemini 3.8 LiveGemini 3.8 Live Extended Thinking
Both models support real-time spoken interactions for voice assistants, conversational experiences, and applications that respond through audio.
Use either model through the AI SDK's realtime API. Install the Gateway provider and a WebSocket client:
Mint a short-lived token, open the WebSocket, and use the model adapter to serialize and parse realtime events:
See the for more details on realtime events and WebSocket connections.realtime quickstart
Try or in the model playground.Gemini 3.8 LiveGemini 3.8 Live Extended Thinking
AI Gateway provides a unified API for calling models, tracking usage and cost, and configuring retries, failover, and performance optimizations for higher-than-provider uptime.
supports real-time audio, visual grounding, automatic switching across 97 languages, and background tool calls while the conversation continues.
google/gemini-3.8-liveadds multi-step reasoning that runs in parallel with speech, allowing it to acknowledge requests and narrate progress without interrupting the conversation.
google/gemini-3.8-live-extended-thinking
Source: original entry ↗