gen‑ai.news
← Back
Video

Introducing Gemini 3.8 Live with Live Avatar

Google DeepMind has announced Gemini 3.8 Live, a new iteration of its real-time conversational model, alongside a feature called Live Avatar. Where earlier versions of Gemini Live focused on low-latency voice interaction, this release adds a visual layer - a generated on-screen figure that speaks and reacts in sync with the model's audio output, giving the exchange a more face-to-face quality.

Live Avatar appears to be aimed at use cases where a purely audio interface feels impersonal or insufficient - think tutoring, customer-facing assistants, or any context where a visible speaker adds useful social cues. The avatar is driven directly by the model's responses, so its expressions and lip movements are tied to what Gemini is actually saying rather than being a separately scripted animation layer.

Gemini 3.8 itself represents an incremental but meaningful update to the underlying model that powers Live interactions. Real-time AI conversation places particular demands on a model - responses need to arrive quickly enough that pauses don't feel unnatural, and the system must handle interruptions, topic shifts, and overlapping speech gracefully. Improvements in this area tend to be felt more as a smoother experience than as a single dramatic capability jump, which makes them easy to underestimate.

The announcement fits into a broader pattern across the AI industry, where developers are moving beyond text and static voice interfaces toward persistent, visually present agents. Google is competing in this space with offerings from companies including OpenAI, which has shown similar avatar-style interaction concepts. How well Live Avatar performs in varied lighting conditions, with different languages, and at scale in real products will be the practical test of whether the feature holds up outside a controlled demo environment.

Enjoy this story? Get the next one in your inbox.

Twice a week: the most important stories in generative image and video AI, distilled into a 2-minute read.

Free. Unsubscribe any time. No spam, ever.

Your next read

No image
Video

Runway’s WorldPrompt and the Engineering of Real-Time Worlds

Runway's WorldPrompt system powers its Gen World Models 2 (GWM 2), enabling real-time generation of video and audio through persistent context and timed actions. Rather than producing discrete clips, the model maintains a continuous understanding of an environment as it unfolds. The approach marks a notable shift in how world models can be steered interactively.

Gemini 3.8 Live with Live Avatar gives Google’s AI a face
Video

Gemini 3.8 Live with Live Avatar gives Google’s AI a face

Google has updated Gemini Live with an animated avatar that lip-syncs and displays facial expressions in real time during conversations. Called Live Avatar, the feature is currently limited to Gemini Enterprise customers and supports 97 languages without degrading video quality. It marks Google's latest step toward giving its AI assistant a visible, expressive presence.

KPop Demon Hunters: How Sony Pictures Imageworks Used Adobe in Creating the Global Phenomenon
Video

KPop Demon Hunters: How Sony Pictures Imageworks Used Adobe in Creating the Global Phenomenon

Sony Pictures Imageworks texture and motion graphics artists detail how Adobe Substance 3D and After Effects formed the backbone of the production pipeline for KPop Demon Hunters, the 2026 Netflix animated film nominated for both Academy Award and Golden Globe honors. From generating over a thousand crowd character variations to crafting intricate Korean embroidery textures, the tools shaped both the scale and the fine detail of the film's distinctive look. Two key artists walk through the speci