🎙 Google Releases Gemini 3.8 Live and Live Extended Thinking
On September 15, Google released live dialogue models: they respond in real time with voice to a stream of audio, video, and text. Extended Thinking reasons and speaks simultaneously, while tool and API calls are performed in the background, without pauses.
🌍 The voice agent no longer "goes silent" during a task: tool calls run in the background, and reasoning happens in parallel with speech. Pricing of $0.75/$4.50 per 1 million text tokens and $3/$12 for audio sets the benchmark for voice agent economics.
👤 You can try it for free in Google AI Studio, and build an agent via the Gemini API. A new voice mode will appear in Gemini Live and Search Live, and in Workspace — Docs Live, Gmail Live, and Keep Live; the models support 97 languages.
Source 1: https://blog.google/innovation-and-ai/models-and-research/gemini-models/gemini-3-8-live-gemini-3-8-live-extended-thinking/ Source 2: https://9to5google.com/2026/09/15/gemini-3-8-live-announced/
