Enterprise & Work

Google's Live Avatar Puts a Human Face on Gemini Enterprise Agents

Google's Live Avatar gives Gemini Enterprise AI agents lip-synced video faces that run API calls mid-conversation. Privacy questions remain unanswered as the feature rolls out to businesses.

Gemini’s Live Avatar Puts a Face on Its AI Agent. It’s Freaking Me Out
Gemini’s Live Avatar Puts a Face on Its AI Agent. It’s Freaking Me Outgwire / Openverse
By James Calloway5 min read

Updated

Why it matters

  • Live Avatar is now available to Gemini Enterprise customers, launched days after Google announced Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking.
  • The feature generates live lip-synced video avatars that process audio, screen sharing, and video feeds simultaneously while running tools and API calls mid-conversation.
  • Google did not respond to questions on general availability or data safeguards for sensitive info like names, addresses, and live video; all audio and video carry SynthID watermarks and custom avatars require verification.

Google has launched Live Avatar, a feature that puts a lip-synced, talking face on its Gemini AI agents for business customers, and the company did so just days after announcing Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking. The rollout signals where Google thinks enterprise AI is heading: away from chat windows and toward AI agents that customers interact with over live video.

Live Avatar is now available to organizations with Gemini Enterprise, according to Google's announcement. Businesses can pick from a small set of prebuilt avatars — cartoon or realistic-looking — and attach one to an AI agent's voice. The result is an agent a customer sees and hears, not just reads.

Google's own example of what this looks like in practice: filing an insurance claim through a video chat with an AI agent. That example is doing a lot of work. It positions Live Avatar squarely in customer support, one of the most aggressive adoption areas for AI agents broadly, and one where the personal stakes for the person on the other end of the call are unusually high.

What the feature actually does

On a technical level, Live Avatar combines an interactive visual experience with a conversational AI agent. It generates live video avatars with lip synchronization, so the avatar's mouth moves in time with the agent's synthesized voice. It processes live audio, screen sharing, and video feeds simultaneously — meaning a customer can share their screen or camera feed with the agent while talking to it.

The agent is not just a talking head. Live Avatar can run tools and execute API calls while continuing its conversation with the customer. In the insurance claim scenario, that means the same avatar that is chatting with a claimant could, in principle, be pulling records or filing data in the background without pausing the interaction.

Live Avatar also introduces native speech-to-speech processing, which Google says produces "more natural interruption recovery without dropping conversation context or backend transactions." In plain terms: if you cut the agent off mid-sentence, it should pick up where things left off without losing the thread of the conversation or whatever backend operation it was running at the time. Interruption handling has been a persistent weak point for voice-based AI systems, and Google is framing this as a core capability rather than a nicety.

The timing ties Live Avatar to Google's broader live-model push. The feature arrives days after the company announced Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking — model variants built specifically for real-time, conversational use rather than one-shot text generation.

The uncomfortable part

The technology works, and that is precisely what makes it unsettling. A customer who calls about an insurance claim may now be looking at a realistic or cartoonish face whose lips move in sync with a synthetic voice while it processes their request. The verisimilitude is the product.

That raises questions Google has not yet answered publicly. The company did not immediately respond to a request for comment on when Live Avatar will be available to the general public, or on how it plans to safeguard the information the AI agent receives during video calls. In the insurance example Google itself chose, that information would include a customer's name, address, and live video while filing a claim.

The gap between the capability and the disclosure is the story here. An AI agent that simultaneously sees your face, hears your voice, reads your shared screen, and triggers API calls against backend systems is collecting and processing an enormous amount of sensitive personal data in real time. Google's announcement describes the interaction model in detail. It says far less about the data handling around it.

Guardrails, of a sort

Google has put some safeguards in place. Businesses cannot deploy an arbitrary face. They choose from curated, prebuilt avatars, according to the company's announcement. Organizations that want a custom avatar must go through a verification process first — a friction point clearly designed to stop bad actors from impersonating specific people at scale.

All audio and video produced by Live Avatar is watermarked with SynthID, Google's system for marking AI-generated content. The watermarking is meant to keep AI-generated audio and video verifiable — a meaningful control as synthetic media becomes harder to distinguish from real recordings by eye or ear.

These measures address some misuse scenarios. Curated avatars and SynthID watermarking make it harder to deploy Live Avatar for covert impersonation, and easier to prove afterward that a given clip came from the system. They do not, by themselves, answer the privacy questions raised by piping live video of claimants into an enterprise AI pipeline, and Google has not publicly detailed how it handles that data.

Why it matters

The commercial logic is straightforward. Businesses are already deploying AI agents to process customer support requests at growing rates, and Live Avatar gives those agents a face, a voice, and a persistent on-screen presence. For enterprises, the pitch is continuity: one agent that talks, watches, shares screens, and executes backend actions without handing the customer off.

The societal question is whether customers will know — or care — that they are talking to a machine, and what consent and disclosure should look like when the machine has a human face. Google's curated-avatar and watermarking requirements are an implicit acknowledgment that the uncanny surface of this technology needs managing. The company's silence on data safeguards for the video calls themselves, and on any timeline for general availability, leaves the harder questions open.

For now, Live Avatar is an enterprise product with enterprise controls. Whether the face of customer service becomes an avatar for everyone will depend on answers Google has not yet provided.

Source: CNET AI

Share this article:

More from James Calloway

James Calloway

Show full bio

News editor covering industry trends and analytics at AI In Context.

118 articles

Related articles

  1. Google Gives Gemini 3.8 Live Agents Talking Avatars
  2. Google Rolls Out Gemini 3.8 Live with Live Avatar for Enterprise
  3. Google Ships Gemini 3.8 Live and Extended Thinking Models
  4. Google Ships Upgraded Gemini 2.5 Flash Native Audio and Live Translation
  5. OpenAI launches Presence, an enterprise agent platform that resolves 75% of its own support calls

« Previous articleNext article »