Wan 3.0 🧠 Model Open source
Alibaba's video family: open-weight Wan 2.x and the 30-second Wan 3.0 API
- GitHub stars
- 18k
- Stars this week
- –
- Forks
- 2.3k
- Licence
- Apache-2.0
- Last push
- 2026-09-21
- Maintainer
- Alibaba (Tongyi Wan)
Where you can use Wan 3.0
Apps in this directory that let you generate with Wan 3.0, from their own documentation.
Good for
About Wan 3.0
What it is
Wan is Alibaba's video generation family. It has two tracks. Earlier versions such as Wan 2.1 and 2.2 were published as open weights under Apache-2.0, and they became the default open video models in ComfyUI. The newest version, Wan 3.0, entered beta on Alibaba Cloud Model Studio in August 2026 and is offered as an API. We found no open weights for Wan 3.0.
Key features
- Wan 3.0: up to 30 seconds per clip, double the 15 seconds of Wan 2.7
- Accepts text, images, video, audio and even documents (PDF, PPT) as input
- Multilingual voices with synchronised facial expressions
- Instruction-based editing of existing clips
- Wan 2.2 (open): 480p/720p text-to-video, image-to-video, speech-to-video and character animation
Where you can use it
Wan 3.0 is available on Alibaba Cloud Model Studio and Qwen Cloud, and in ComfyUI partner nodes (Wan 3.0 and 2.7). Wan 2.6/2.7 is documented in Higgsfield, Krea, Leonardo.Ai and invideo AI. The open Wan 2.x weights run locally in ComfyUI.
Pricing and rights
Wan 2.1/2.2 weights are free to use under Apache-2.0, including commercial use. You need your own GPU. Alibaba had not published Wan 3.0 API pricing at the beta announcement. In partner apps, Wan uses that app's credits.
Who it is for
Tinkerers and studios that want a capable open model to run and fine-tune locally. Also developers who want a long-clip API from a major cloud.
Verdict
Wan is the most practical open video family, and the Apache licence is creator-friendly. The best new features (Wan 3.0, editing, multi-reference) are API-only. Details like resolution and pricing for 3.0 were still thin in beta.
Pros
- Apache-2.0 open weights for Wan 2.1/2.2
- Wan 3.0 makes 30-second clips with voices
- Huge ComfyUI community (LoRAs, workflows)
- Documents and web pages accepted as input
Cons
- Wan 3.0 is API-only (no open weights found)
- Wan 3.0 pricing and resolution not published at beta
- Local runs need a strong GPU
Similar AI models
All video models →Veo 3.1 🧠 ModelPaid
Google DeepMind's text- and image-to-video model with native sound, up to 4K
Kling 3.0 🧠 ModelFreemium
Kuaishou's video model with native multilingual audio, lip sync and 15-second clips
Seedance 2.5 🧠 ModelPaid
ByteDance's video model: 30-second clips with audio, many references and local edits
MiniMax H3 (Hailuo 3.0) 🧠 ModelFreemium
MiniMax's open-weight video model: native 2K, stereo sound, 15-second clips
Runway Gen-4.5 🧠 ModelFreemium
Runway's own video model, now with native audio and one-minute multi-shot scenes
LTX-2.5 🧠 ModelOpen source
Lightricks' open-weight video model with synchronised audio and 4K output