gen‑ai.news
← Back
Video

Runway wants to turn AI video generation into a live stream you control in real time

Runway wants to turn AI video generation into a live stream you control in real time

Runway is exploring a shift in how AI video generation is delivered - moving away from the familiar model of submitting a prompt and waiting for a completed clip, toward something closer to a live stream that responds to user input as it unfolds. The idea is that generation and playback would happen simultaneously, giving users a more interactive, continuous experience rather than a series of discrete renders.

The technical foundation for this is GWM-1, Runway's generative world model, which produces video frame by frame rather than all at once. That sequential architecture makes real-time streaming more feasible, since the system does not need to compute an entire clip before anything can be shown. The model maintains a sense of continuity across frames, which is essential if the output is to feel coherent when a user steers it mid-stream with a new prompt or instruction.

The practical implications for creative work could be significant. Video production workflows today involve a lot of iteration - generating a clip, reviewing it, adjusting the prompt, and generating again. A streaming model that responds to input in real time would compress that loop considerably, letting creators shape output on the fly rather than waiting through multiple render cycles. It also opens up interaction patterns that feel less like file generation and more like a controllable, generative environment.

Runway also points to industrial and scientific use cases beyond creative tooling. Robotics and autonomous driving both depend heavily on simulation - systems need vast amounts of varied, realistic visual data to train on, and generating that data quickly and interactively has real value. A real-time world model that can be steered by prompts or conditions could serve as a flexible simulation environment, producing training scenarios on demand. Whether Runway brings this capability to market as a standalone product or integrates it into its existing platform remains to be seen, but the direction signals that the company is thinking about video generation as something closer to a live, responsive medium than a rendering pipeline.

Enjoy this story? Get the next one in your inbox.

Twice a week: the most important stories in generative image and video AI, distilled into a 2-minute read.

Free. Unsubscribe any time. No spam, ever.

Your next read

No image
Video

Runway’s WorldPrompt and the Engineering of Real-Time Worlds

Runway's WorldPrompt system powers its Gen World Models 2 (GWM 2), enabling real-time generation of video and audio through persistent context and timed actions. Rather than producing discrete clips, the model maintains a continuous understanding of an environment as it unfolds. The approach marks a notable shift in how world models can be steered interactively.

Gemini 3.8 Live with Live Avatar gives Google’s AI a face
Video

Gemini 3.8 Live with Live Avatar gives Google’s AI a face

Google has updated Gemini Live with an animated avatar that lip-syncs and displays facial expressions in real time during conversations. Called Live Avatar, the feature is currently limited to Gemini Enterprise customers and supports 97 languages without degrading video quality. It marks Google's latest step toward giving its AI assistant a visible, expressive presence.

No image
Video

Introducing Gemini 3.8 Live with Live Avatar

Google DeepMind has introduced Gemini 3.8 Live, an updated multimodal model paired with a new Live Avatar feature that generates an animated, talking on-screen presence during real-time conversations. The combination allows users to interact with a responsive visual agent rather than a purely voice-based interface. The release marks another step in Google's effort to make AI interactions feel more immediate and embodied.