Technology

Google's AI Assistant Now Has a Face That Talks to You

Martin HollowayPublished 7d ago3 min readBased on 8 sources
Reading level
Google's AI Assistant Now Has a Face That Talks to You
source:blog.google

Google has released a talking face for its AI assistant for business customers. The feature, called Gemini 3.8 Live with Live Avatar, is now generally available to Gemini Enterprise customers. The company explained it on September 24, 2026, in a post titled 'Introducing Gemini 3.8 Live with Live Avatar' Google. The idea is simple. You talk with the AI and watch it talk back.

The avatar is live video of a talking face, matched to a computer-made voice from the gemini-3.8-live model Documentation. It runs through the Gemini Live API, a connection that keeps voice and video delays very short. For businesses, the building tools sit in the Gemini Enterprise Agent Platform, a system to build, manage and control AI assistants.

The face moves like a person on a call. It moves its lips in time with speech and changes expression during conversations The Verge. Google says it also moves its head naturally. People can take turns speaking without long pauses, and the system sees visual input almost at once. It feels more like a video call than sending messages back and forth.

Google describes Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking as its most advanced models so far for natural conversation. Live Avatar is the visual add-on to those models. It does not replace text or voice answers. It adds a matching face to them.

For now, only Gemini Enterprise customers can use it. Google offers it with US and EU endpoints and with reserved capacity Google Cloud. Endpoints set where the system runs, which helps meet rules about keeping data in a certain region. Reserved capacity means businesses book space in advance, which keeps delays steady even when many people use it at once.

It works in 97 languages and can switch languages without the video breaking or the face slipping out of sync. Google will offer ready-made avatars and let companies make their own. In practice, a support helper, sales assistant, training coach or internal helper can look the same in every country without rebuilding the video system for each language.

Each video carries a hidden mark. Live Avatar output includes an invisible SynthID watermark, a signal you cannot see. It stays with the video and is meant to show later that the video was made by AI.

The broader picture for teams building these helpers is that a small change can help a lot. A face that moves with speech gives clues about timing and pauses that voice alone does not give. Keeping the face steady when languages change also avoids a common problem, where the voice switches fine but the face freezes or lags.

In my view, starting with businesses is on purpose. Fun avatars for shoppers get quick attention. Business avatars get tested. Call handling, task completion, handoffs to people, and user trust can be compared with text-only and voice-only versions. Google says Live Avatar lets enterprises expand their virtual offerings, and early use will likely be in customer-facing jobs with high volume and clear steps, where one steady persona lowers cost while people can still step in.

For the people who will run this system, the practical point is that live video needs a strong connection. It is sensitive to uneven delays and lost data. Reserved capacity and regional endpoints help, but teams will still need to check call quality, backup plans, and the hidden mark. The face makes the system easier to follow. It also makes problems easier to spot.

The longer history here gives reason for hope. We moved from typed commands to windows to touch screens to voice. Each step made it easier to say what we want. We have seen this pattern before with home computers, when simpler controls brought in more users. A face that listens, speaks and remembers context keeps moving that way. It will help work when it is managed well, tested honestly, and used where seeing someone helps people get things done.