changelog

Google Adds Real-Time Video Avatars to Gemini 3.8 Live for Enterprise Voice Agents

TL;DR

Google has added Live Avatar to its Gemini 3.8 Live dialogue models, pairing real-time video personas with voice AI for enterprise customer service and sales use cases. The feature is restricted to Gemini Enterprise customers via allowlisting and includes SynthID watermarking on all generated video.

2 min read
0

Google has rolled out Live Avatar, a new feature that adds real-time video personas to its Gemini 3.8 Live and Live Extended Thinking dialogue models, according to a company blog post. The addition follows last week's launch of Gemini 3.8 Live, which Google described as its most advanced live dialogue system for building voice agents.

Live Avatar pairs near real-time visual presence with Gemini's existing voice capabilities, creating what Google calls "an experience that listens, sees and speaks with a dynamic visual persona." The company's blog demonstrates both realistic and cartoon-style avatars performing lip-syncing, facial expressions, and what Google terms "fluid turn-taking" during conversation.

The feature is aimed squarely at enterprise customer service and sales applications. Because the avatars run on Gemini's underlying reasoning models, Google says they can trigger tool calls and retrieve data in the background while maintaining an active spoken dialogue with a user — for example, pulling account information mid-conversation without breaking the visual interaction.

Availability and customization

Google will provide a library of preset avatars for organizations to deploy out of the box. Enterprises can also build custom avatars from reference images to match "brand styling or character identity," though this customization tier is gated behind enterprise allowlisting — it is not available to general developers or consumer accounts.

Every avatar output is watermarked using Google's SynthID system, which the company says is intended to keep AI-generated video content identifiable and address concerns about synthetic media transparency. Google describes this as part of "strict safeguards designed to respect identity" for the feature.

Access restrictions

Live Avatar is currently limited to Gemini Enterprise customers. Google has not disclosed pricing for the feature, nor has it published specific latency, resolution, or benchmark figures for how the avatars perform compared to competing video-agent products from other vendors.

No technical specifications — frame rate, resolution, supported languages, or per-minute cost — were included in Google's announcement. The company has also not stated whether Live Avatar will expand beyond enterprise allowlisting to broader developer access on a defined timeline.

What this means

This is a feature layered on top of the existing Gemini 3.8 Live model family, not a new model release — the underlying dialogue and reasoning models are unchanged. Google is packaging its live voice AI into a more commercially deployable form factor by adding a face to it, targeting a specific enterprise use case: replacing or augmenting human customer service and sales reps with video-capable agents.

The SynthID watermarking and identity safeguards suggest Google is anticipating backlash over synthetic human-like avatars appearing in customer-facing roles, a concern reflected in early public reaction to demo videos showing photorealistic AI personas. Whether businesses adopt Live Avatar at scale will likely hinge less on the technology's technical fluency and more on whether customers tolerate — or actively reject — interacting with an AI face during support and sales interactions. Google has not published data on user acceptance, only its own product claims.

Related Articles

changelog

Cline v4.1.21 Fixes Local Model Timeouts, Expands Catalog to 6,386 Models Across 209 Providers

Cline's v4.1.21 update adds a new OpenAI-compatible provider called ai&, refreshes its model catalog to 6,386 models across 209 providers, and fixes a bug that ended tasks prematurely when local models hit output-token limits. Eleven unpinned providers, including GitHub Copilot and Vertex, now default to Claude Opus 5.5.

changelog

OpenAI Python SDK v3.19.0 Adds GCP Storage Support and References Unreleased 'GPT-Rosalind' Model

OpenAI released v3.19.0 of its Python SDK on September 22, 2026, adding GCP external storage support and a code reference to an unannounced research model called GPT-Rosalind. The release also ships five bug fixes covering WebSocket handling, retry logic, and async compatibility.

changelog

OpenAI Rolls Out Improved Prompt Caching for GPT-6

OpenAI has updated its prompt caching system for GPT-6, adding explicit cache breakpoints, new diagnostic tools, and finer-grained controls. The company claims the changes improve cache hit rates and reduce both latency and cost for repeated-context API calls.

changelog

OpenAI Cuts GPT-6 Sol and Luna Prices in Half, but Independent Benchmarks Show Flat Performance

OpenAI's GPT-6 Sol and Luna cut input/output token prices in half versus GPT-5.6, with Sol now at $2/$10 per million tokens and Luna at $0.10/$0.50. Independent testing from Artificial Analysis shows intelligence scores barely moved, with regressions on some knowledge-work benchmarks.

Comments

Loading...