Google DeepMind publishes introduction of Gemini 3.8 Live with Live Avatar enterprise feature
Google DeepMind Launches Gemini 3.8 Live with Real-Time Video Avatars
The feature extends conversational AI into real-time animated video personas at enterprise scale, covering 97 languages. Google states SynthID watermarking is embedded in all output to keep AI-generated content detectable and reduce misinformation.
The full picture
Google DeepMind released Gemini 3.8 Live with a Live Avatar feature, an enterprise capability that pairs near real-time video generation with speech to produce animated AI personas capable of lip-synced, expressive conversation. The feature supports simultaneous vision and audio input, asynchronous tool calling for background data fetching without interrupting conversation, and native speech-to-speech synchronization across 97 languages without degrading video quality. Organizations can create custom avatars from a high-quality reference image to preserve brand styling or character identity, though this requires enterprise allowlisting. All audio and video output is watermarked with SynthID.
How it developed
Sources
Related
- Continues inAI agent hacking incidents across labs
- Grew out ofGoogle DeepMind Gemini 3.8 Live voice models
- Continues inGoogle DeepMind's Gemini 3.8 Flash TTS
Want this in your inbox?
I send one email each morning with the stories that moved. If you would rather just read here, that works too.
Subscribe free