Lifestyle

Google launches Gemini 3.8 Live with Live Avatar: a talking, expressive AI avatar for businesses

On September 24, 2026, Google announced Gemini 3.8 Live with Live Avatar, which adds near-real-time generated avatar video to live voice conversations, available first in Gemini Enterprise. This article summarizes the features, limitations and what it means for ordinary users, based on official Google materials.

About 5 min read

Google launches Gemini 3.8 Live with Live Avatar: a talking, expressive AI avatar for businesses
Image: Mokaair (Original editorial artwork)

What Google announced

On September 24, 2026, Google published Gemini 3.8 Live with Live Avatar on its official blog. Google said this is a new feature following last week's release of Gemini 3.8 Live, combining near-real-time video generation with voice so the AI can converse with people as an animated avatar that listens, watches and speaks.

According to Google, Live Avatar offers precise lip sync, natural facial expressions and smooth turn-taking, and can be used in business scenarios such as customer service or interactive guides. Google said the feature is available in Gemini Enterprise from launch day, with API documentation for developers (an API is the interface developers use to connect their own software to Google's AI).

Google launches Gemini 3.8 Live with Live Avatar: a talking, expressive AI avatar for businesses
Mokaair editorial verification flow · Image: Mokaair (Original editorial artwork)
Read the full description

Sources are collected, independently checked, then reviewed by Jev.

Key features: what Google says

  • Multimodal input (handling several kinds of input at once): Google says Live Avatar processes visual and audio input simultaneously, then responds with expressive voice and video.
  • Asynchronous tool calling (doing lookups in the background without pausing the conversation): Google says the avatar can call tools and retrieve data while the conversation continues; the demo scenario is a hotel guest check-in.
  • Multilingual: Google says it can switch among 97 languages and claims this does not reduce video fidelity or cause visual drift, meaning the avatar's appearance does not gradually degrade or shift.
  • Custom avatars: beyond the preset avatar library, developers can generate custom avatars from high-quality reference images; Google says this is currently available only through an enterprise allowlist, meaning only business customers Google has approved.

Gemini 3.8 Live vs. Live Avatar

Sources: Google blog and Google DeepMind Gemini 3.8 Audio model card. Tokens are the small units of data an AI model processes.
ItemGemini 3.8 LiveWith Live Avatar
Output typesAudio and textAudio, video and text
Output limit64K tokens24K tokens
AvailabilityGemini API, Gemini App, Google AI Studio, Gemini Enterprise Agent Platform, Google Search Live (per model card)Gemini Enterprise (per Google blog)
Continuous interaction timeNot separately stated in model cardMinutes, not hours (per model card)

According to the Google DeepMind model card, Gemini 3.8 Audio is based on Gemini 3 Pro, and Gemini 3.8 Live's input can include audio, images, video and text, with a context window (the amount of material the model can take in at once) of up to 128K tokens. The model card also mentions general foundation-model limitations such as possible hallucinations (confidently stating things that are not true), occasional slowness or timeouts, and a knowledge cutoff of January 2025.

Safety and transparency: SynthID watermarks

Google says all outputs generated by its AI products carry a SynthID watermark embedded directly in the audio and video. According to Google DeepMind, SynthID watermarks are imperceptible to humans but detectable by SynthID technology; users can upload an image, video or audio to Gemini and ask whether it was generated or modified by Google AI. Google DeepMind also says the SynthID Detector verification portal is currently being tested with journalists and media professionals.

The model card further states that, based on evaluation results for Gemini 3.7 Flash, Google considers models such as Gemini 3.8 Live unlikely to reach the tracked or critical capability levels in its Frontier Safety Framework, Google's own system for flagging potentially dangerous AI abilities. This is Google's own assessment, not the conclusion of an independent external review.

FAQ

Can ordinary people use Live Avatar now?

According to the Google blog, Gemini 3.8 Live with Live Avatar is currently offered in Gemini Enterprise, mainly for businesses and developers; custom avatars are further restricted to an enterprise allowlist. Ordinary users are more likely to encounter it indirectly through a company's customer service or guide services.

What languages can Live Avatar speak?

Google says Live Avatar can switch among 97 languages and adjusts lip sync and expressions when switching languages mid-conversation. Google's announcement does not list the full set of languages.

Can I talk with it for a long time?

No. The Google DeepMind model card states that the version with Live Avatar supports minutes, not hours, of continuous interaction; it may also occasionally be slow or time out.

How can I tell the person on screen is AI-generated?

Google says outputs from its AI products carry an embedded SynthID watermark. According to Google DeepMind, you can upload an image, video or audio to Gemini and ask whether it was generated or modified by Google AI. This method only identifies Google AI content.

Have these feature claims been independently verified?

No. All information in this article comes from official Google and Google DeepMind pages, including the features, the number of languages and the safety assessment; all are Google's own claims.

Browse the latest news in this topic

Latest travel guides

Sources

Lifestyle