AlibabaAlibaba·🎬 Video Generation

Wan 3.0

anonymized
Try on Venice.ai ↗
Quick reference
Wan 3.0 — TLDR
  • 🆕 Alibaba's newest Wan video model on this catalog, August 2026
  • 📏 Alibaba states up to 30 seconds of video per generation
  • 🌐 Positioned by Alibaba as generating video from any input
  • 🔧 This entry is the text-to-video path of the Wan 3.0 release
  • 🎯 Photorealistic frames with strong subject coherence across shots
  • 🏢 Delivered through Alibaba Cloud Model Studio video generation APIs
  • 👁️ Sibling variants cover image-to-video and reference-to-video workflows
  • 📚 No public parameter count, license, or weight release listed
💰 Pricing
$0.060 – $3.30
per generation
📅 On Venice since
Aug 1, 2026
23 days ago
Provider

Alibaba Group is a Chinese multinational technology company founded in 1999 and headquartered in Hangzhou, Zhejiang. Originally built around e-commerce and cloud computing, Alibaba has become one of the most prolific contributors to open-weight AI research,…

Read full profile →
68 models on Venice
29 video · 22 text · 7 image · 6 inpaint · 2 embedding · 2 tts
Since Jan 11, 2025

About this model

Wan 3.0 is the text-to-video entry point into the newest generation of Alibaba's Wan visual model line, which is developed alongside — but separate from — the Qwen language models such as Qwen 3.8 27B. On this catalog it is described as a photorealistic generator with strong subject coherence across frames, and it is dated August 2026.

The headline change versus earlier releases in this family, Wan 2.7 and Wan 2.6, is clip length and input breadth: Alibaba presents Wan 3.0 as producing up to 30 seconds of video in a single generation, and as accepting a wide range of input types rather than prompts alone. Longer single-pass output matters because it reduces the need to stitch several short clips together to build a continuous shot.

Access is through Alibaba Cloud Model Studio, whose video generation service exposes text-driven and image-driven synthesis endpoints. Alibaba does not publish a parameter count or weights for this release, so it is used as a hosted service rather than run locally.

Companion Wan 3.0 variants on this catalog handle the other input paths: Wan 3.0 image-to-video animates a supplied still, while Wan 3.0 Reference conditions generation on reference material. For downloadable Wan models, the catalog also lists Wan 2.2 Enhanced.

This About section is AI-generated from public sources (Claude Opus 5), with no human editing. It may contain inaccuracies — verify critical details against the sources listed above.

Data sources: Venice API · HuggingFace · Wikipedia — enrichment updated 12h ago