🗣️ Introducing Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking. These advanced audio models are built for natural conversation, featuring major upgrades in turn-taking and near real-time reasoning. They also significantly streamline how you build intelligent voice agents. Until now, advanced reasoning and reliability came at the cost of high latency and the complexity of cascaded, multi-model pipelines. Stitching together separate speech-to-text, reasoning, and text-to-speech models adds delay and inflates costs. By replacing that stack with a single multimodal API call, these two models make building highly responsive voice agents simpler and cost-effective. 🧵 Here’s what you can build with them: