Introducing Gemini 3.8 Live with Live Avatar

2026-09-27 · Google DeepMind

Introducing Gemini 3.8 Live with Live Avatar

Building on the momentum of last week's Gemini 3.8 Live launch, on September 24, 2026, Google introduced Gemini 3.8 Live with Live Avatar. This new feature brings near real-time visual presence to native live dialogue models. By pairing near real-time video generation with speech, the Live Avatar feature creates an experience that listens, sees, and speaks with a dynamic visual persona.

With precise lip-syncing, natural expressions, and fluid turn-taking, Live Avatar enables enterprises to expand their virtual offerings more interactively. Whether providing engaging customer service or delivering interactive walkthroughs, it transforms digital exchanges into richer, more accessible experiences. Starting today, Gemini 3.8 Live with Live Avatar is available in Gemini Enterprise.

More Natural and Multimodal Conversations

Conversation is inherently multimodal: we listen, look, speak, and use facial expressions to communicate. Live Avatar brings these capabilities to enterprise agents.

  • Simultaneous Input Processing: By processing visual and audio inputs simultaneously, it generates enriching conversations for a more comprehensive experience.
  • Near Real-Time Response: It takes in what it sees and hears in near real time, responding with expressive audio and video for a more natural conversation.

Asynchronous Tool Execution with Continuous Presence

Beyond visual presence, the feature is backed by Gemini’s advanced reasoning capabilities.

  • Background Tool Calling: Live Avatar can trigger tool calls and fetch data in the background while continuing active dialogue.
  • Seamless Complex Task Handling: For example, when checking in a guest at a hotel, it can call tools in the background while the dialogue continues uninterrupted, ensuring a smooth conversational flow.

Conversational Experiences Built for Global Scale

Conversational presence should feel natural and not be limited by languages.

  • Native Multilingual Synchronization: Live Avatar features native multilingual speech-to-speech synchronization.
  • Seamless Language Transition: The feature dynamically adapts its lip-sync and expressions, seamlessly transitioning across 97 languages without degrading video fidelity or introducing visual drift.

A Live Avatar to Fit Your Brand Needs

Organizations often need distinct visual identities to fit their brand.

  • Preset Avatar Library: In addition to a library of diverse, preset avatars, organizations can customize their Live Avatars.
  • Customization Options: From a high-quality reference image, developers can generate a fully animated, responsive avatar while preserving reference likeness, brand styling, or character identity.
  • Availability Note: Custom avatar creation is currently available only through enterprise allowlisting.

Trust and Transparency at Its Core

Live Avatar was built with strict safeguards designed to respect identity and keep AI-generated content transparent.

  • SynthID Watermarking: All output generated by AI products is watermarked with SynthID. This imperceptible watermark is woven directly into the audio and video output.
  • Misinformation Prevention: This helps ensure AI-generated content remains detectable to minimize misinformation and misattribution.
  • Safety and Responsibility: Users can read the model card to explore the comprehensive approach to safety and responsible deployment.

Get Started

Gemini 3.8 Live with Live Avatar is available in Gemini Enterprise. Developers can explore the API documentation to get started.

Source