Google launches Gemini 3.8 Live models that can reason while they talk
Google launched Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking models, which continue reasoning and calling tools in the background while a voice response is in progress. The models are available via Gemini API and AI Studio, support 97 languages, with audio input at $3 per million tokens and audio output at $12.
- Extended Thinking scores 82.6 on Speech-to-Speech Quality Index
- Agentic tasks completed in 68.6% of cases per Artificial Analysis
- Audio input $3 per million tokens (~$0.005/min), output $12 (~$0.018/min)
- Models switch between 97 languages in a single conversation
Read next
AI