gen‑ai.news
← Back
Video

YouTube Shorts Gets AI Remix Feature Powered by Gemini Omni

YouTube Shorts Gets AI Remix Feature Powered by Gemini Omni

YouTube has rolled out a remix capability for Shorts that uses Gemini Omni to let viewers transform videos they encounter on the platform. Accessed through a remix icon at the bottom of any eligible Short, the feature offers a "reimagine" option where users can prompt the model to apply a visual style - pixel art, anime, found-footage horror - or alter the actual content of the clip, adding background elements, changing costumes, or inserting the user into the scene.

The self-insertion capability is the most technically significant aspect. Users can place themselves into another person's Short, which requires the model to handle both style transfer and compositional editing at the same time. All output is watermarked with SynthID, connecting it to Google's broader provenance infrastructure announced at I/O.

Creator controls are built in: uploaders can toggle whether their videos are eligible for reimagining. This matters particularly for creators who post footage of children or content they do not want stylistically altered or repurposed, though the opt-out mechanism places the burden on creators rather than defaulting to protection.

The feature is a concrete consumer-facing application of Gemini Omni's video capabilities, and it introduces interactive AI video editing to a platform with enormous reach. How YouTube moderates the output - particularly given the potential for the self-insertion tool to be used in misleading or harassing ways - will be worth watching as the feature scales.

Read at The Verge →
Share:X

Enjoy this story? Get the next one in your inbox.

Twice a week: the most important stories in generative image and video AI, distilled into a 2-minute read.

Free. Unsubscribe any time. No spam, ever.

Your next read

No image
Video

Runway’s WorldPrompt and the Engineering of Real-Time Worlds

Runway's WorldPrompt system powers its Gen World Models 2 (GWM 2), enabling real-time generation of video and audio through persistent context and timed actions. Rather than producing discrete clips, the model maintains a continuous understanding of an environment as it unfolds. The approach marks a notable shift in how world models can be steered interactively.

Gemini 3.8 Live with Live Avatar gives Google’s AI a face
Video

Gemini 3.8 Live with Live Avatar gives Google’s AI a face

Google has updated Gemini Live with an animated avatar that lip-syncs and displays facial expressions in real time during conversations. Called Live Avatar, the feature is currently limited to Gemini Enterprise customers and supports 97 languages without degrading video quality. It marks Google's latest step toward giving its AI assistant a visible, expressive presence.

No image
Video

Introducing Gemini 3.8 Live with Live Avatar

Google DeepMind has introduced Gemini 3.8 Live, an updated multimodal model paired with a new Live Avatar feature that generates an animated, talking on-screen presence during real-time conversations. The combination allows users to interact with a responsive visual agent rather than a purely voice-based interface. The release marks another step in Google's effort to make AI interactions feel more immediate and embodied.