gen‑ai.news
← Back
Video

Google AI Releases Gemini Omni 1.1 Flash: 40-Second Scene Extension, First/Last Frame Control, and 4K Upscaling

Google has shipped Gemini Omni 1.1 Flash as a production update to its multimodal video generation and editing model. The release focuses on three concrete capability areas rather than a broad architectural overhaul, and the changes address some of the more practical limitations users encountered with the previous version.

The most significant shift is in how scene extension works. Previously, the model used only the final frame of an existing clip as context when generating a continuation. With 1.1 Flash, it can now analyze up to 10 seconds of prior footage before extending the scene. That wider temporal window helps the model maintain consistent lighting, motion direction, and subject positioning across the join point - issues that often created jarring cuts in earlier outputs. The maximum extension length also reaches 40 seconds in this release.

Frame-level control has been added through first and last frame pinning. Users can now specify both the opening and closing frames of a generated clip, giving them a more direct way to choreograph camera movement and scene composition without relying entirely on text prompts. A related addition allows existing video clips to be passed in as reference material for character consistency, which is useful for productions that need a recurring subject to look the same across multiple generated segments.

The 4K upscaling feature completes the update, letting lower-resolution outputs be brought up to higher display standards in a post-generation step. Gemini Omni 1.1 Flash sits within Google's broader Gemini model family, which has been positioning native video understanding and generation as a core capability rather than a separate pipeline. These targeted additions suggest Google is iterating toward more precise editorial control, moving the tool closer to something that fits into professional or semi-professional video workflows.

Enjoy this story? Get the next one in your inbox.

Twice a week: the most important stories in generative image and video AI, distilled into a 2-minute read.

Free. Unsubscribe any time. No spam, ever.

Your next read

AI-generated videos are already displacing actors and livestreamers across China's entertainment industry
Video

AI-generated videos are already displacing actors and livestreamers across China's entertainment industry

China's short-drama industry has moved rapidly toward AI-generated content, with 95 percent of the 128,000 short dramas released in Q1 2026 produced using AI. Actors report being asked to surrender their voice and likeness data before losing their jobs, and labor disputes tied to AI displacement are climbing. The trend offers an early, concrete look at how generative video is reshaping entertainment workforces at scale.

LAION drops massive open video dataset with 10 million hours of footage for AI research
Video

LAION drops massive open video dataset with 10 million hours of footage for AI research

LAION has released its Big Video Dataset (BVD), a large open collection of 80 million videos totaling 10 million hours of footage, aimed at advancing AI video research. The dataset also includes 55 million automatically described clips, and models trained on it have outperformed those trained on the previous benchmark dataset, InternVid. Its open availability marks a notable moment for researchers who have lacked large-scale, freely accessible video data.