Google Launches Gemini 3.8 Live Featuring Interactive Digital Avatars That Support 97 LanguagesAI
26 Sept 2026, 9:50 am (18 min ago)· 0

Google Launches Gemini 3.8 Live Featuring Interactive Digital Avatars That Support 97 Languages

Google has rolled out Gemini 3.8 Live with Live Avatar, an upgrade allowing enterprise users to interact in real time with human-like avatars that can see, hear, and speak.

Real-time artificial intelligence interaction is moving beyond simple text and voice. Google has introduced Gemini 3.8 Live with Live Avatar, an upgrade designed to bring a visual face to its Gemini AI platform. This development means businesses can now deploy interactive digital assistants capable of hearing, seeing, and speaking back almost instantaneously.

A More Human-Like Interaction Experience

The introduction of the Live Avatar feature represents a significant shift from traditional voice-only AI responses. The digital characters are designed to mimic natural human behavior, complete with realistic facial expressions and lip movements that are synchronized directly with the output voice. This technology aims to make interactions feel more authentic, reducing the artificial barrier between users and software.

Also read

Key Visual and Audio Features for Businesses

Tailored primarily for commercial applications, this tool is currently accessible through the Gemini Enterprise tier. Companies can utilize these life-like avatars for diverse roles, including advanced customer service desks, virtual assistants, and live product demonstration tools.

What sets Gemini 3.8 Live with Live Avatar apart is its ability to handle multiple streams of data at once. The AI can process audio and visual inputs simultaneously. This means it can listen to a user's spoken words while also reading what is displayed on a screen, formulating a coherent response that incorporates both inputs. The final reply is delivered using synced voice, video, and appropriate facial expressions.

Asynchronous Capabilities and Broad Language Integration

To maintain a seamless flow of conversation, Google has integrated asynchronous tool calling. In practice, this means the AI does not need to pause the dialogue when fetching data or processing tasks. It can continue talking with the user in the foreground while executing background operations, such as checking a database or running an analytical tool.

Communication barriers are also addressed through massive language support. The digital avatar is capable of interacting in 97 different languages. When a user switches languages mid-conversation, the avatar instantly adapts, modifying its pronunciation, facial expressions, and lip movements to align perfectly with the phonetics of the chosen tongue.

Customization Options and Security Measures

Organizations using this system have access to several pre-designed avatars and a variety of voice profiles. Furthermore, businesses can generate their own custom digital avatars. By feeding high-quality reference images into the platform, organizations can create animated characters that match their brand's visual identity.

Currently, the option to build customized avatars is limited to select enterprise clients. To ensure safety and transparent deployment, Google has embedded SynthID watermarking technology. This system places a digital watermark on the avatar videos, making it straightforward for viewers to recognize that the content has been generated by artificial intelligence.

Questions & Answers

What is Gemini 3.8 Live with Live Avatar?
It is a new visual AI feature from Google that allows users to interact in real time with an animated digital avatar that can see, hear, and speak.
How many languages does the Live Avatar support?
The feature supports 97 different languages, adapting its pronunciation, lip-syncing, and facial expressions accordingly.
Can companies create their own custom avatars?
Yes, companies can generate custom avatars using high-quality reference images, though this feature is currently limited to select enterprise clients.
How does asynchronous tool calling work in Gemini 3.8 Live?
It allows the AI to perform background tasks, like retrieving information, while maintaining a continuous and uninterrupted conversation with the user.
What is SynthID, and why is it used here?
SynthID is Google's watermarking technology used to place a digital watermark on AI-generated avatar content, making it easy to identify.

Comments 2

Rohan Gupta@rohan-gupta·just now

Lip-syncing across 97 languages sounds amazing, but will it actually catch all the different regional accents and local dialects properly?

Ravikash Gupta@ravikash·just now

Rohan, that's the real test! 97 languages sound huge, but let's see how well it catches our varied local accents and regional tones.

Citizen journalism

Become a TrendKia journalist

Voice of the people

Share news, photos and videos from your area with TrendKia and let your voice reach the nation. Every citizen a journalist.

Join now
CH 01 LIVE
TrendKia TV ON AIR
Chamar no WhatsApp