Gemini 3.8 Live Adds Real-Time Animated Persona for Enterprise Conversations

Google DeepMind has introduced Gemini 3.8 Live with Live Avatar, a feature that adds a real-time animated visual persona to its conversational AI. The system synchronizes lip movements and facial expressions with speech, and can process visual and audio inputs together. Available in Gemini Enterprise, it also supports asynchronous tool calls to handle tasks while maintaining dialogue.
The feature builds directly on last week's Gemini 3.8 Live launch, adding a video-generation layer to the existing live dialogue system. It processes audio and visual inputs together, enabling the avatar to respond with synchronized speech, lip movements, and facial expressions in near real time. The system also runs asynchronous tool calls in the background, allowing it to fetch data or complete tasks while maintaining an uninterrupted conversation.
The avatar supports native speech-to-speech synchronization across 97 languages, adapting lip-sync and expressions without visual degradation. Organizations can choose from preset avatars or generate custom ones from reference images, though custom creation requires enterprise allowlisting. All output carries SynthID watermarks embedded in both audio and video to maintain content transparency.
This technology could reshape customer-facing industries by making AI agents feel more human and engaging, potentially improving user satisfaction in sectors like hospitality, retail, and support services. However, realistic animated personas may also blur the line between human and machine interaction, raising concerns about user trust and emotional attachment. The built-in watermarking suggests an awareness of misuse risks, but the broader societal effects—particularly around employment displacement and the normalization of AI-mediated conversations—remain uncertain.