
Alibaba released its new AI video model, Wan3.0, in beta. It can create videos up to 30 seconds long from text, PDFs, and PowerPoint files.
The company aims to fix visual drift and distortion in AI videos.
It is priced per second, from $0.05 to $0.28.
What happened
Alibaba's video generation model Wan3.0 is now available in beta. It produces videos up to 30 seconds long and accepts text, PDFs, web pages, and PowerPoint files as input.
Why it matters
This doubles the video length compared to its predecessor, Wan2.5. The model also recommends the best video length based on the user's prompt and includes an extension tool to make existing videos longer. Wan3.0 processes text, images, video, and audio at the same time.
What to watch
It is available through the wan.video website, Alibaba Cloud Model Studio, or via API on Qwen Cloud. There are two tiers: a Standard version, currently at a 30 percent discount, and a faster Prime version. Alibaba is pitching Wan3.0 for uses from film production to robotics training.
Ask the AI about this article →
The launch of Wan3.0 comes as Alibaba ramps up its AI spending significantly. The company just announced the largest share sale by a Hong Kong-listed company to fund its AI push, and last week reported a 75 percent year-over-year drop in quarterly profit driven by sharply higher AI investments. This new model is a key part of that strategy, targeting a wide range of commercial applications.
Alibaba is positioning Wan3.0 for professional and business use cases. These range from speeding up film production and generating short dramas to creating marketing and training videos from text and images. The company is also courting tech developers who could use it to produce realistic simulation footage for training autonomous vehicles and robotics systems, areas where consistent visual details are crucial.
The model's focus on maintaining consistency in characters, props, and spatial layouts addresses a known weakness in AI-generated video, which often suffers from visual drift and distortion. By improving on this and supporting more complex inputs like multi-page documents, Alibaba is attempting to make its model a more practical tool for converting static business data into dynamic video content.
AI-summarized, only the topics you pick — one digest a day via Email, Slack, or Discord.
Free · takes 30 seconds · unsubscribe anytime
Ask AI anything about this article. Q&As are published on this page for other readers too.
Oneiric, an AI-generated open source video project, has been released

AMD posted record revenue of $11.54 billion with Data Center revenue of $6.72 billion, up 107% year over year

Higgsfield AI, a startup, produced The Cully Hill Boys, a 110-minute action-comedy featuring licensed celebrit…

A developer created pagedMark, a tool designed to remove AI provenance markers—both visible labels and invisib…

A post on LessWrong poses the question of whether a large language model (LLM)—an AI system that understands a…

OpenAI closed an $852 billion self-funded tender offer on August 10, 2026, then lost chief operating officer B…
