Latest
Google Adds Reasoning to Its Real-Time Voice AI, Opening the Door to Smarter Phone Agents
Google DeepMind released Gemini 3.8 Live and Live Extended Thinking, adding a reasoning mode to its real-time voice/video API. Live agents built on it can now pause to work through multi-step requests — like troubleshooting or scheduling with conditions — before responding, instead of relying on scripted, single-turn flows.
What changes for operators — For a 10-200 person B2B company running phone support or a sales qualification line through voice AI, this closes the biggest gap in current voice agents: handling anything beyond a single-turn lookup. A support call that requires checking an order status, then applying a conditional refund rule, then confirming with the customer previously needed a handoff to a human or a scripted decision tree. Extended Thinking lets the agent reason through that sequence live, on the call, which means fewer escalations and shorter average handle time for the tier-one queue. Teams evaluating or already running voice bots for inbound support or outbound qualification should treat this as the point to re-test latency and accuracy on their actual call scripts — reasoning modes typically add processing time, so the tradeoff between depth and response speed needs to be measured before rolling it into a live queue, not assumed.