gen‑ai.news
← Back
Video

Black Forest Labs makes FLUX 3 Video generally available and claims it beats Seedance 2.0

Black Forest Labs makes FLUX 3 Video generally available and claims it beats Seedance 2.0

Black Forest Labs has made FLUX 3 Video generally available, bringing its latest video generation model to a wider audience after what appears to have been a limited early access period. The model produces Full HD clips with a maximum duration of 20 seconds - a meaningful step up in length and resolution compared to many competing offerings that still cap out at shorter runtimes or lower resolutions.

One of the more technically notable additions is native audio generation paired with lip-synced dialogue. Rather than requiring a separate audio layer to be composited onto generated footage, FLUX 3 Video handles both simultaneously, and the lip-sync capability works across more than 14 languages. That kind of multilingual support at the generation stage - rather than as a post-processing dubbing step - reduces friction for creators working across different regional markets.

The model also introduces in-scene typography rendering, allowing readable text to appear naturally within generated video frames. Legible text in AI-generated images and video has historically been a weak point for diffusion-based models, so its inclusion as a feature worth highlighting suggests BFL has made targeted improvements in that area.

On the competitive positioning side, Black Forest Labs is citing its own Elo-based ranking system to place FLUX 3 Video above Google's Gemini Omni Flash and ByteDance's Seedance 2.0. Elo rankings derived from user preference votes can be a useful signal, but it is worth noting that BFL is both running and reporting these benchmarks, which warrants some caution in taking the comparisons at face value. Independent third-party evaluations will ultimately offer a clearer picture. Black Forest Labs, the team behind the widely used FLUX image generation models, has been working to expand into video, and this release marks a concrete step in that direction.

Enjoy this story? Get the next one in your inbox.

Twice a week: the most important stories in generative image and video AI, distilled into a 2-minute read.

Free. Unsubscribe any time. No spam, ever.

Your next read

No image
Video

Runway’s WorldPrompt and the Engineering of Real-Time Worlds

Runway's WorldPrompt system powers its Gen World Models 2 (GWM 2), enabling real-time generation of video and audio through persistent context and timed actions. Rather than producing discrete clips, the model maintains a continuous understanding of an environment as it unfolds. The approach marks a notable shift in how world models can be steered interactively.

Gemini 3.8 Live with Live Avatar gives Google’s AI a face
Video

Gemini 3.8 Live with Live Avatar gives Google’s AI a face

Google has updated Gemini Live with an animated avatar that lip-syncs and displays facial expressions in real time during conversations. Called Live Avatar, the feature is currently limited to Gemini Enterprise customers and supports 97 languages without degrading video quality. It marks Google's latest step toward giving its AI assistant a visible, expressive presence.

No image
Video

Introducing Gemini 3.8 Live with Live Avatar

Google DeepMind has introduced Gemini 3.8 Live, an updated multimodal model paired with a new Live Avatar feature that generates an animated, talking on-screen presence during real-time conversations. The combination allows users to interact with a responsive visual agent rather than a purely voice-based interface. The release marks another step in Google's effort to make AI interactions feel more immediate and embodied.