Google Introduces Gemini 3.8 Live Audio Models for Real-Time Voice AI Workflows

Chronological Source Flow
Back

AI Fusion Summary

Google DeepMind has launched Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking, advanced audio models for real-time voice AI workflows. These models support 97 languages, process live visual inputs, and execute API calls during conversations. Extended Thinking leads the Speech to Speech Quality Index. Available via Gemini API and Google AI Studio at $0.005/min, both models include SynthID watermarking. They enable businesses to implement production-grade voice agents for conversational task handling and customer-facing interactions.
Community Comments
Loading updates...
0