gen‑ai.news
← Back
Video

Vidu Q4 AI Video Model with Native Audio

Vidu Q4 AI Video Model with Native Audio

Shengshu has launched Vidu Q4, a new iteration of its AI video generation model that introduces native audio as a core capability rather than a post-processing add-on. Earlier versions of Vidu focused primarily on visual output, so the integration of audio directly into the generation pipeline marks a meaningful shift in scope for the model.

Native audio in video generation means the model can produce synchronized sound - whether ambient noise, effects, or other audio elements - as part of the same generation pass that produces the video frames. This approach can result in better temporal alignment between what is seen and what is heard, compared to workflows that stitch audio onto video after the fact using separate models.

Beyond audio, Vidu Q4 also brings improvements to motion quality and prompt control. Stronger prompt adherence is a recurring focus across the video generation field, as models have historically struggled to faithfully translate detailed text descriptions into consistent on-screen motion and composition. Better motion quality, meanwhile, suggests refinements to how the model handles movement over time - reducing common artifacts like unnatural warping or inconsistent subject motion across frames.

Vidu Q4 is available through the Vidu platform at vidu.com. Shengshu, the Beijing-based AI company behind Vidu, has been developing the model as a competitive offering in a market that includes tools from Runway, Kling, and others. The addition of native audio gives Vidu Q4 a feature that remains relatively uncommon among publicly available video generation models, making it a notable option for creators looking to produce video content with cohesive sound from a single generation tool.

Read at shengshu →
Share:X

Enjoy this story? Get the next one in your inbox.

Twice a week: the most important stories in generative image and video AI, distilled into a 2-minute read.

Free. Unsubscribe any time. No spam, ever.

Your next read

Her AI-Generated Video Swayed the Judge. The Court Said it Carried 'Undue Emotional Weight'
Video

Her AI-Generated Video Swayed the Judge. The Court Said it Carried 'Undue Emotional Weight'

An AI-generated video of a murder victim was used as a victim impact statement in an Arizona court, with the deceased man's sister creating an avatar to speak in his place. While the judge was visibly moved by the presentation, an appeals court later found that the video carried "undue emotional weight" - raising serious questions about the role of generative AI in legal proceedings.

No image
Video

Reka Releases Rho-1: A 19B Omni-Reasoning Model That Understands, Generates Video and Outputs Robot Actions in One

Reka has released Rho-1, a 19-billion-parameter model that handles text, images, video understanding, video generation, and robot action outputs within a single unified network. The architecture uses a shared key-value cache across all modalities, and a distilled variant can produce a 5.3-second video clip in roughly one second. The release is currently a research preview, with no public weights available.

BlackRAW Studio App Lets You Color Grade and Edit ProRes RAW on iPhone
Video

BlackRAW Studio App Lets You Color Grade and Edit ProRes RAW on iPhone

BlackRAW Studio is a new iPhone app designed for filmmakers who need to review, color correct, and edit Blackmagic RAW and ProRes RAW footage directly on their device. It brings a level of RAW video handling to mobile that has traditionally required a desktop workstation. The app targets professionals who want a capable on-set or on-the-go editing option without reaching for a laptop.