New Models
AI Minute Newsroom
2026-08-07
Alibaba's Wan 3.0 enters public beta: 30-second video from your documents, slides and webpages
Alibaba's Tongyi Lab put its Wan 3.0 video model into public beta on Wednesday. The new version generates videos up to 30 seconds with sound, and folds what used to be separate models — text-to-video, image-to-video, reference-based generation and editing — into one system. The headline trick: it accepts documents, spreadsheets, presentations and webpages as input, turning static, text-heavy material into finished video.
Why it mattersThirty seconds with native audio moves AI video past the clip-toy stage, but the document-to-video angle is the bigger shift: it aims the tool at offices, classrooms and e-commerce rather than filmmakers. Alibaba's Wan family also underpins much of the open video ecosystem, so what lands in Wan usually spreads well beyond Alibaba.
✓ Verified · 3 sources
Read in the app — free, in 9 languages
Related stories
Reflection will hand out a 501-billion-parameter model for free.
2026-10-06Microsoft's new transcriber starts writing before you finish speaking.
2026-10-05An American lab is about to give away a China-class open model.
2026-10-05A video site now gives away the best-scoring open translation model.
2026-10-05Half the people on a short video call thought the AI was human.
2026-10-04