Why it matters
This advancement allows developers to build more engaging and natural multimodal conversational agents. The ability to process visual and audio inputs simultaneously, coupled with asynchronous tool execution and multilingual support, opens new avenues for creating richer, more accessible digital interactions for businesses.

What changed

Google DeepMind has launched Gemini 3.8 Live with Live Avatar, an extension of its Gemini 3.8 Live conversational AI. This new feature introduces near real-time visual presence, allowing Gemini to respond with a dynamic visual persona that includes speech and video generation. The Live Avatar aims to create a more natural and intuitive conversational experience by pairing near real-time video generation with speech, enabling precise lip-syncing, natural expressions, and fluid turn-taking. This capability is now available within Gemini Enterprise.

The feature enhances multimodal conversations by processing visual and audio inputs simultaneously, generating richer dialogue. It supports asynchronous tool execution, allowing the AI to trigger tool calls and fetch data in the background while maintaining an active dialogue, ensuring an uninterrupted conversational flow for complex tasks. Furthermore, Live Avatar offers native multilingual speech-to-speech synchronization, adapting lip-sync and expressions seamlessly across 97 languages without degrading video fidelity. Organizations can also customize avatars using a high-quality reference image to match brand identity, though custom avatar creation is currently available via enterprise allowlisting.

Trust and transparency are addressed through watermarking all AI-generated output with SynthID, an imperceptible watermark embedded in the audio and video to ensure detectability and minimize misinformation. A model card detailing the approach to safety and responsible deployment is also available.

Why it matters for builders

For AI builders, Gemini 3.8 Live with Live Avatar provides tools to create more immersive and interactive enterprise applications. The integration of visual presence, alongside advanced reasoning and asynchronous tool handling, allows for the development of sophisticated agents capable of managing complex tasks while maintaining engaging user interactions. The extensive multilingual support and customizable avatar options offer flexibility for global deployments and brand alignment.

Practical impact

Developers can leverage Gemini 3.8 Live with Live Avatar within Gemini Enterprise to build next-generation customer service bots, interactive walkthroughs, and virtual assistants. The ability to generate responsive avatars that match brand aesthetics and communicate fluently in multiple languages can significantly enhance user engagement and accessibility. Builders should explore the API documentation to integrate these advanced visual and conversational capabilities into their applications, focusing on use cases that benefit from a more human-like AI presence.

Caveats and source limits

The source material indicates that custom avatar creation is currently available only through enterprise allowlisting. While the feature supports 97 languages, specific details on performance benchmarks or independent evaluations of its real-time capabilities and lip-sync accuracy across all supported languages are not provided. The exact pricing for Gemini Enterprise, which includes this feature, is also not detailed in the announcement.

Sources

Written with AI assistance from the linked sources; every claim below was checked against them automatically. How we produce articles.

Claim check: 10/10 supported claims - 10 evidence links - 100% avg confidence
  • Gemini 3.8 Live with Live Avatar integrates real-time visual presence into Gemini's conversational AI.supported - deepmind.google
  • Live Avatar enables a more natural and intuitive conversational experience by coupling live dialogue capabilities with low-latency streaming video.supported - deepmind.google
  • The feature creates an experience that listens, sees, and speaks with a dynamic visual persona, featuring precise lip-syncing, natural expressions, and fluid turn-taking.supported - deepmind.google
  • Gemini 3.8 Live with Live Avatar is available in Gemini Enterprise.supported - deepmind.google
  • The feature processes visual and audio inputs simultaneously to generate richer conversations.supported - deepmind.google
  • Live Avatar supports asynchronous tool execution, allowing it to trigger tool calls and fetch data in the background while continuing active dialogue.supported - deepmind.google
  • Live Avatar features native multilingual speech-to-speech synchronization, adapting to 97 languages without degrading video fidelity.supported - deepmind.google
  • Organizations can customize Live Avatars from a reference image to preserve likeness and brand styling.supported - deepmind.google
  • Custom avatar creation is currently available only through enterprise allowlisting.supported - deepmind.google
  • All AI-generated output is watermarked with SynthID for transparency and detectability.supported - deepmind.google

Caveats

  • Single-source caution: verify critical details at the linked source.
Radar score 74/100 - how it was calculated
Reliability90
Freshness50
Novelty67
Technical53
Developer60
Ecosystem86
Confidence100
  • Reliability 90: Primary official source
  • Freshness 50: Fresh official source date
  • Novelty 67: Official announcement
  • Technical 53: Structured technical source signals
  • Developer 60: Builder relevance source signals
  • Ecosystem 86: Official source
  • Confidence 100: Claims have reliable evidence
Share
XLinkedInHacker News

Related articles

Enterprise AI - Jun 4, 2026Endava Uses AI Agents and ChatGPT Enterprise for Software DeliveryEndava is leveraging AI agents, ChatGPT Enterprise, and Codex to enhance software delivery processes. The company aims to automate workflows and foster an AI-native culture within its enterprise.Developer Tools - Jun 2, 2026GitHub Copilot Adds Gemini 3.1 Pro and 3.5 Flash ModelsGitHub Copilot now supports Gemini 3.1 Pro (Preview) and Gemini 3.5 Flash across its CLI, cloud agent, and Copilot app. Access to these models varies by Copilot subscription tier.Enterprise AI - Sep 25, 2026Google Private AI Compute Adds Secure Server-Side MemoryGoogle's Private AI Compute platform is introducing a new private, server-side memory layer. This feature aims to provide persistent, cross-device AI memory while maintaining on-device privacy standards.Other - Jul 29, 2026Google Cloud Platform's Generative AI Repository: A Resource for Gemini and Vertex AI DevelopmentThe GoogleCloudPlatform/generative-ai repository provides sample code and notebooks for developing Generative AI applications on Google Cloud, specifically leveraging Gemini Enterprise Agent Platform. It serves as a practical resource for builders working with large language models and Google Cloud's AI services like Vertex AI. The repository has accumulated 17,516 stars and 4,385 forks, indicating significant community interest.Infrastructure - Aug 4, 2026FastAgent: Deploying Local AI Agents as Live ServicesFastAgent is a TypeScript project designed to transform local AI agent directories into live, accessible services. It supports deployment across various channels including in-app integrations, GitHub, and Telegram.Other - Oct 3, 2026Graph-It-Live VS Code Extension v1.16.1 ReleasedThe Graph-It-Live VS Code Extension, a TypeScript-based tool for visualizing dependencies and AI agent interactions, has released version v1.16.1. This update follows a recent release and contributes to the project's AI and developer signal counts.