Google introduces interactive visual avatars for Gemini
Agents & ToolsThere's An AI For That · 1h ago

Google introduces interactive visual avatars for Gemini

Google launched Live Avatar for enterprise Gemini customers, adding synchronized real time video faces to customer service interactions. The avatars support 97 languages, adjust facial expressions dynamically, and execute administrative queries in the background while conversing.

Google

The Blend

Google has launched a new feature called Live Avatar for its Gemini Enterprise platform. The update adds dynamic, video-based faces to automated voice interactions, letting corporate customer service bots speak, react, and display realistic facial expressions while conversing with clients.

According to Google's official announcement, the digital faces can automatically adjust their lip movements and expressions across 97 different languages. The system is also designed to execute administrative background tasks, such as pulling up hotel reservation details or processing database queries, without pausing or disrupting the ongoing video call.

For everyday consumers, this technology means future interactions with corporate support desks could feel much closer to a video call with a human representative than typing into a text widget. Having a visual face that responds in real time could make automated assistance feel less robotic and easier to navigate for non-technical users.

However, replacing human staff with highly realistic digital personas raises important questions about user trust and acceptance. It remains to be seen whether consumers will embrace talking to a computer-generated face or if the visual presentation will trigger an uncomfortable uncanny valley effect during routine support calls.

Written independently by AI News Smoothie from the reporting listed below. Facts belong to the original publishers. Follow the links for their full coverage.

Ingredients

  • Introducing Gemini 3.8 Live with Live Avatar

    Google launched a live video avatar capability for enterprise users that combines real-time speech dialogue with synchronized facial animations across nearly one hundred languages.

Read the original