Google has released Gemini 3.8 Live and 3.8 Live Extended Thinking with asynchronous tool execution capabilities.
Google has launched Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking, bringing asynchronous tool execution and advanced background reasoning to real-time, low-latency conversational models. These models allow complex, multi-step background tasks to be processed while simultaneously maintaining a continuous live audio stream with the user.
The ability to execute tools asynchronously while holding a live conversation is a major leap forward for agentic voice assistants.
- –This solves the latency and "awkward silence" problems when a voice agent needs to fetch data or execute a slow API call.
- –Extended Thinking enables the model to perform highly complex reasoning tasks seamlessly in the background.
- –Opens up robust enterprise possibilities for real-time customer support, inventory management, and interactive troubleshooting.
DISCOVERED
3h ago
2026-09-19
PUBLISHED
17h ago
2026-09-18
RELEVANCE
AUTHOR
octavusai