Google DeepMind and Google AI Studio introduced two live dialogue audio models: Gemini 3.8 Live (scale, speed, cost efficiency) and Gemini 3.8 Live Extended Thinking (higher-complexity multi-step reasoning). Both add near real-time visual understanding, automatic detection across 97 languages, and background tool calling without breaking conversation flow. Developers can build in public preview via AI Studio and the Gemini API; consumer rollouts include Search Live and Gemini App, with enterprise private preview.

Key Takeaways

  • โœ“Two live audio models: 3.8 Live (scale/speed/cost) and 3.8 Live Extended Thinking (complex multi-step reasoning).
  • โœ“Near real-time vision, 97-language auto-detection, and background tool calling without breaking chat flow.
  • โœ“Developers: public preview in AI Studio and Gemini API; consumer Search Live/Gemini App; enterprise private preview.
ADSponsored