Reported by 1 source

The short version

  • Google has launched a real-time animated avatar feature for its Gemini 3.8 Live model, allowing users to see facial expressions and lip movements during conversations.
  • The technology is currently limited to Enterprise customers and supports seamless transitions across ninety-seven languages without compromising video quality.
  • Organizations can select from preset avatars or create custom ones, with all outputs protected by invisible watermarks and identity-respecting safeguards.

Google has expanded the capabilities of its artificial intelligence platform by introducing a visual component to real-time conversations. The update, part of the Gemini 3.8 Live suite, enables users to interact with an animated persona that responds instantly to dialogue. This development marks a shift from purely text or audio-based interactions toward more immersive, human-like interfaces within the company’s enterprise offerings.

The core functionality of this new feature involves precise synchronization between speech and visual cues. The avatar lip-syncs accurately to the spoken words and displays varying facial expressions that align with the tone and content of the conversation. This level of detail aims to make digital interactions feel more natural and engaging, bridging the gap between static text responses and dynamic human communication.

News Journal

Access to this technology is currently restricted. Only customers subscribed to Gemini Enterprise can utilize the Live Avatar feature at this stage. This limitation suggests that Google is prioritizing business applications where enhanced user engagement and professional presentation may offer significant value. The company appears to be testing the waters with a controlled rollout before considering broader availability.

A notable technical achievement of this update is its multilingual capability. The system supports ninety-seven languages, allowing it to switch between them seamlessly during a single conversation. According to Google, these transitions occur without degrading video fidelity or introducing visual drift, which can sometimes plague real-time rendering systems. Demonstrations have shown the avatar speaking in English and Japanese with mouth animations that correctly match the phonetics of each language.

Beyond simple conversation, the avatar can also retrieve and display information on-screen while speaking. This multimodal approach allows the AI to provide visual aids or data points simultaneously with verbal explanations, potentially enhancing clarity and user understanding. The integration of real-time visual input processing, a feature highlighted in the recent Gemini 3.8 Live model release, underpins this ability to handle complex, multi-sensory interactions.

Customization options are available for organizations looking to tailor the experience to their brand or specific needs. While Google provides a library of preset avatars for immediate use, enterprises also have the ability to create their own unique personas. This flexibility allows businesses to maintain brand consistency and control over how their AI representatives appear to customers or employees.

Security and ethical considerations are embedded into the design of this feature. All output generated by the Live Avatar includes an invisible SynthID watermark, a measure intended to help identify AI-generated content. Additionally, Google has implemented safeguards designed to respect identity, ensuring that the technology is used responsibly and does not misrepresent individuals or violate privacy norms.

The introduction of real-time animated avatars represents a significant step in the evolution of human-computer interaction. By combining advanced language processing with high-fidelity visual rendering, Google is pushing the boundaries of what AI assistants can do. As this technology matures and potentially expands beyond enterprise users, it could redefine expectations for digital communication tools.

Looking ahead, the success of this feature will likely depend on user adoption and the practical benefits it delivers in professional settings. If enterprises find that animated avatars improve engagement or clarity, Google may consider making similar features available to consumer users. For now, the focus remains on refining the technology within a controlled environment.

This update underscores the ongoing competition among tech giants to create more intuitive and immersive AI experiences. As other companies develop their own visual AI interfaces, the race to provide the most realistic and useful digital personas intensifies. Google’s move with Gemini 3.8 Live positions it as a leader in this emerging space, setting new standards for real-time AI interaction.

Sources behind this briefing

Go to the original reporting

  • The Verge↗Gemini 3.8 Live with Live Avatar gives Google’s AI a face