gen‑ai.news
← Back
Video

Alibaba's Wan3.0 generates AI videos up to 30 seconds long from text, images, and documents

Alibaba's Wan3.0 generates AI videos up to 30 seconds long from text, images, and documents

Alibaba has launched Wan3.0, the latest version of its video generation model, expanding both the length and input flexibility of AI-generated video clips. Where many competing tools cap output at a few seconds or require a single prompt type, Wan3.0 accepts text, images, PDF documents, and PowerPoint files as source material - broadening the range of practical use cases, particularly for business and presentation-oriented content.

The model supports video generation at up to 1080p resolution and can produce clips as long as 30 seconds, a duration that remains relatively uncommon among publicly available video generation systems. Alibaba has set the price at $6 per 30-second 1080p clip, positioning it as a commercial offering rather than a free research preview. The ability to ingest structured document formats like PDFs and slide decks is a notable differentiator, potentially making it easier for users to convert existing materials into video without manual reformatting.

Wan3.0 builds on Alibaba's ongoing investment in generative AI, a push that has come with significant financial costs. The company reported a 75 percent drop in quarterly profit compared to the same period a year earlier, a decline directly tied to increased AI-related capital expenditure. This pattern - trading near-term earnings for long-term AI positioning - mirrors spending trends seen at other large technology companies investing heavily in model development and infrastructure.

The broader context here is a competitive landscape in AI video generation that has grown considerably more crowded over the past year, with offerings from companies like OpenAI, Google, Runway, and others all targeting different segments of the market. Wan3.0's document-input capability and relatively long clip duration give it a distinct angle, though real-world output quality and consistency will ultimately determine how it fares against established alternatives.

Enjoy this story? Get the next one in your inbox.

Twice a week: the most important stories in generative image and video AI, distilled into a 2-minute read.

Free. Unsubscribe any time. No spam, ever.

Your next read

World models that ignore human beliefs predict the wrong actions, new research shows
Video

World models that ignore human beliefs predict the wrong actions, new research shows

Current AI world models simulate physical environments but leave out a critical layer: what people believe, want, and intend. New research introduces a "Mental World Modeling" framework that adds these mental variables, and finds that even smaller models using it can outperform larger ones that ignore human psychology. The key bottleneck turns out to be modeling how physical and mental states evolve together over time.

Runway News | The Next Phase of Enterprise Video Generation
Video

Runway News | The Next Phase of Enterprise Video Generation

Runway's Chief Revenue Officer has distilled hundreds of enterprise conversations into five themes shaping how large organizations are approaching AI video generation. The piece covers everything from model consolidation and data sovereignty to shifting cost structures and the move toward autonomous execution.

Major YouTube creators are facing backlash for accepting AI money
Video

Major YouTube creators are facing backlash for accepting AI money

Several prominent filmmaking YouTubers, including Matti Haapoja and Sam "Kold" Kolder, have drawn criticism after posting sponsored content promoting Higgsfield's AI video platform without clearly disclosing the paid nature of those partnerships. The backlash intensified when other creators began sharing apparent screenshots of outreach from PR firms working on Higgsfield's behalf. The episode has sparked a broader conversation about transparency and trust in the creator community around AI tool