❯ Alibaba Cloud Video Generation Model Wan3.0 Enters Public Beta, Generates 30 Seconds per Run and Supports Document Input
BETAAlibaba Cloud’s next-generation video generation model Wan3.0 has opened public beta, generating up to 30 seconds of video in a single run. On top of the four base modalities of text, image, audio, and video, it supports document-format input for the first time, including doc, xls, ppt, pdf, and md. The beta spans Alibaba Cloud Bailian, Wanjing Yike, the Wanxiang official website, and Qwen Creation on PC, with the Qwen App in staged gray release.
PRICINGAPI pricing is split into three tiers by resolution: 480P, 720P, and 1080P at 0.3 yuan, 0.6 yuan, and 1.2 yuan per second, respectively, and the API will be fully opened in the near term. A caveat worth flagging: these are beta-period reference prices only; official pricing has yet to be announced. The 0.6-yuan-per-second rate at 720P is a meaningful threshold for teams mass-producing content such as short dramas and ad clips—the model cost for a finished 30-second video lands at around 18 yuan.
USE CASESDocument input deserves more attention than the duration number. Teams producing enterprise content can now drop a PPT or a financial-statement spreadsheet straight into the model and get a finished video out, removing the manual middle step of “first rewriting the document into storyboard prompts.” The input side of video models thus expands from prompts to structured files, pulling users a big stride from the creative side toward the enterprise-office side—and directly reshaping the labor-cost structure of content teams.
▪ SIGNALA video model that can chew through PPTs and spreadsheets isn’t just taking the editor’s job anymore.