AI · Sep 24, 2026
Gemini 3.8 Live with Live Avatar gives Google’s AI a faceIntroducing Gemini 3.8 Live with Live Avatar
Sep 24, 2026, 9:20 AM · Google DeepMind

Google DeepMind pairs Gemini 3.8 Live with streaming video avatars—lip-sync, 97 languages, async tools—shipped first to Gemini Enterprise with SynthID watermarks.
Why it matters
Conversation is multimodal. DeepMind’s Live Avatar couples near-real-time video generation with live speech so enterprise agents can listen, see, and speak with a visual persona.
It is available in Gemini Enterprise starting September 24, 2026, with preset characters, custom avatars from a reference image (allowlisted), and background tool calls that keep dialogue moving.
This is the enterprise face of the live-model wave: customer service and walkthroughs that look like a person, not a typing bubble.
From the desk
I’m watching the product bet, not the demo reel. Live Avatar claims precise lip-sync, natural expressions, and fluid turn-taking, plus native multilingual speech-to-speech across 97 languages without visual drift. For global support desks, that is the difference between a novelty and a staffing plan.
The useful-AI case is straightforward: hotels, retail, and education want agents that stay present while tools fetch data. Asynchronous tool execution while the avatar keeps talking is the right architecture for that.
The harm path is equally clear. A branded, lip-synced face lowers the bar for trust—and for deception. DeepMind puts SynthID watermarks in audio and video and points to a model card. That is necessary. It is not sufficient if enterprises deploy these faces without clear disclosure to the human on the other end.
Custom avatars from a single reference image, even behind allowlisting, raise likeness and brand-abuse questions the moment the allowlist leaks or a partner ships carelessly. We’re for richer, more accessible digital service. We’re not for pretending the avatar is a person.
Context
The post builds on the prior week’s Gemini 3.8 Live launch. Authors Shuo-yiin Chang and CJ Zheng wrote on behalf of the Gemini Audio Team.
Who feels it
- Enterprises
- Avatar agents can cover multilingual support and walkthroughs, but disclosure and identity policy need to ship with the feature.
- Developers
- Enterprise API path matters more than consumer demos; custom avatar allowlisting will gate early experiments.
- Regulators / platforms
- Watermarked synthetic faces still need UI-level labeling when they talk to customers.
What to watch
- How Gemini Enterprise customers label Live Avatar sessions to end users
- Abuse reports involving custom reference-image avatars
- Latency and fidelity reviews once the feature leaves launch demos
Companies: Google