gen‑ai.news
← Back
Video

xAI adds character references and 1080p to Imagine Video 1.5

xAI adds character references and 1080p to Imagine Video 1.5

xAI has rolled out Imagine Video 1.5, a notable update to its AI video generation model that introduces several new input and output capabilities. The headline additions are image-based character references and voice references, which allow users to anchor the appearance and sound of characters in generated video to specific source material rather than relying entirely on text descriptions.

The update also introduces a prompt-only mode, meaning users who do not have reference images or audio can still generate video from text alone - keeping the workflow accessible while the reference-based features serve those who need tighter creative control. Multi-reference support extends this further, letting users feed in more than one reference simultaneously to guide a single generation, which is useful for scenes involving multiple characters or blended visual styles.

On the output side, Imagine Video 1.5 now renders natively at 1080p. Previous generations of AI video tools have often topped out at lower resolutions, so native full-HD output is a meaningful step for users who want footage that holds up in production or on larger screens without upscaling artifacts.

xAI's Imagine suite - which also includes its image generation tool - sits within the broader Grok ecosystem the company has been building. Pushing Imagine Video toward more reference-driven, higher-resolution output positions it closer to competing tools from providers like Runway, Kling, and others that have made character consistency a key selling point. Whether the reference fidelity holds up under varied prompts and complex scenes will be the practical test for creators considering it as part of their workflow.

Enjoy this story? Get the next one in your inbox.

Twice a week: the most important stories in generative image and video AI, distilled into a 2-minute read.

Free. Unsubscribe any time. No spam, ever.

Your next read

Watching Roku’s AI channel is like eating from a trough
Video

Watching Roku’s AI channel is like eating from a trough

Roku has launched a 24/7 free ad-supported streaming channel dedicated entirely to AI-generated content, sourced from a startup called Fairground. The move marks one of the more visible attempts to bring generative video into mainstream living-room viewing. Whether audiences will warm to it is another question.

See what 5 builders are making with Gemini Omni
Video

See what 5 builders are making with Gemini Omni

Google's Gemini Omni lets users generate and edit video through natural conversation, and a new spotlight from the company shows how five independent builders are putting that capability to practical use. The examples range from visualizing abstract ideas to streamlining video editing workflows. Together, they offer a concrete look at how conversational video AI fits into real creative and production work.

Black Forest Labs makes FLUX 3 Video generally available and claims it beats Seedance 2.0
Video

Black Forest Labs makes FLUX 3 Video generally available and claims it beats Seedance 2.0

Black Forest Labs has moved FLUX 3 Video out of early access and into general availability, offering Full HD video generation with clips up to 20 seconds long. The model includes native audio output and lip-synced dialogue across more than 14 languages. According to BFL's own Elo benchmark rankings, it outperforms both Gemini Omni Flash and Seedance 2.0.