Google DeepMind publishes details of Gemini 3.8 Live with Live Avatar, including SynthID watermarking and 97-language support
Google DeepMind Launches Gemini 3.8 Live with Animated AI Avatars
The feature brings lip-synced, multilingual AI video personas to enterprise use cases, combining conversational AI with real-time video generation. Google DeepMind states that SynthID watermarking is applied to all output to help minimize misinformation and misattribution.
The full picture
Google DeepMind has introduced Gemini 3.8 Live with Live Avatar, an enterprise feature that pairs near real-time video generation with speech to produce animated AI personas capable of lip-synced, expressive conversations. The feature supports simultaneous vision and audio input, asynchronous tool calling for background data fetching during conversations, and native speech-to-speech synchronization across 97 languages without degrading video quality. Organizations can create custom avatars from a reference image to preserve brand styling or character identity, though that capability requires enterprise allowlisting. All audio and video output is watermarked with SynthID to keep AI-generated content detectable.
How it developed
Sources
Related
- Grew out ofAI agent hacking incidents across labs
- Grew out ofOpenAI's GPT-6 Astra model
- Grew out ofGoogle DeepMind Gemini 3.8 Live voice models
- Grew out ofGoogle DeepMind's Gemini 3.8 Flash TTS
Want this in your inbox?
I send one email each morning with the stories that moved. If you would rather just read here, that works too.
Subscribe free