Google has introduced two new advanced audio models, Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking, designed to enhance voice agent capabilities and make AI conversations more natural. Gemini 3.8 Live focuses on speed, cost-efficiency, and handling mid-sentence interruptions across 97 languages, with visual context integration for tasks like step-by-step instructions. Gemini 3.8 Live Extended Thinking offers deeper intelligence for complex, multi-step tasks by reasoning and speaking in parallel, even narrating its progress. AI
IMPACT These advanced audio models are expected to significantly improve voice agent capabilities and make AI interactions more natural and efficient for consumers, developers, and enterprises.
RANK_REASON Frontier-lab model release with system card.
- Gemini 3.8 Live
- Gemini 3.8 Live Extended Thinking
- Gemini API
- Gemini app
- Gemini Enterprise
- Gmail
- Google AI
- Google AI Studio
- Google Docs
- Google Workspace
- Search Live
- Artificial Analysis
- Big Bench Audio
- EVA-Bench
- Fishjam
- Gemini Enterprise Agent Platform
- Google DeepMind
- ServiceNow
- Speech Agent Arena
AI-generated summary · Google Gemini · from 4 sources. How we write summaries →