Google DeepMind Launches Gemini 3.8 Live and 3.8 Live Extended Thinking

Google DeepMind has released Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking, marking a significant update to its real-time multimodal model family. The Live variants are designed for low-latency, streaming audio and video interactions, while the Extended Thinking version adds a reasoning layer that allows the model to deliberate before responding. This combination targets agentic and voice-first applications where both speed and accuracy matter, making it directly relevant to developers building real-time assistants or multi-turn voice agents. The extended thinking capability in particular aligns Gemini more closely with reasoning-focused models, giving builders a new tool for complex task decomposition in live contexts. Developers can access these models through Google AI Studio and the Gemini API.
Read original source ↗Part of the 2026-09-16 briefing→