HeadlinesBriefing favicon HeadlinesBriefing.com

Gemini 3.8 Live: Advanced Voice AI Models

Google DeepMind Blog •
×

Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking are our most advanced live dialogue models yet. Major upgrades in intelligence and parallel reasoning make them more intuitive to collaborate with and use to execute complex tasks using your voice.

Today, we’re introducing two new models that bring advancements in near real-time reasoning to more effectively enable voice agents and make conversing with AI feel more intuitive and intelligent. Gemini 3.8 Live: Built for scale and cost efficiency, combining conversational intelligence with fluid dialogue and visual grounding. Gemini 3.8 Live Extended Thinking: Built for high-complexity tasks, with increased intelligence and multi-step reasoning.

Gemini 3.8 Live Extended Thinking provides enterprise-grade task completion and intelligence, capturing the #1 overall spot on Artificial Analysis' Speech to Speech Quality Index (82.6), and leads in agentic task completion with 68.6% on τ-Voice and 35.1% on Sierra’s τ-Voice-banking benchmark. It also provides strong reasoning capabilities, scoring 97.7% on Big Bench Audio. Gemini 3.8 Live has shown a high preference among users, securing a second place in the Speech Agent Arena.

Gemini 3.8 Live processes visual inputs in near real-time, enriching conversations with context. It automatically detects and transitions between 97 supported languages mid-conversation. For tasks that require deeper reasoning, 3.8 Live Extended Thinking reasons and speaks simultaneously. Across Google Workspace and Search, our Live models deliver more intuitive, collaborative experiences.