OpenAI's voice model doesn't think, and that's the point
Google released Gemini 3.8 Live and Extended Thinking, while OpenAI released GPT-Live-1. Google keeps speech, reasoning and tool calls in one session; OpenAI separates the voice model from a backend reasoning model, leaving orchestration to the developer.
- Gemini 3.8 Live Extended Thinking handles speech, reasoning and tool calls in one session
- GPT-Live-1 handles dialogue while a backend model like GPT-6 Astra does reasoning
- GPT-Live-1 response latency is about 800 ms, priced at $0.05 per voice minute
- Gemini 3.8 Live costs $0.005 per input minute and $0.018 per audio output minute
Read next
AI