Try “Veo”, “voiceover”, “thumbnail” or “Descript” · Esc to close

Wan 3.0 🧠 Model Open source

Alibaba's video family: open-weight Wan 2.x and the 30-second Wan 3.0 API

Video models · Open source ★ 18k · Apache-2.0 · updated 2026-09-21

8.0editor score
Visit Wan 3.0 ↗
GitHub stars
18k
Stars this week
–
Forks
2.3k
Licence
Apache-2.0
Last push
2026-09-21
Maintainer
Alibaba (Tongyi Wan)

Where you can use Wan 3.0

Apps in this directory that let you generate with Wan 3.0, from their own documentation.

See every app × model on the model map →

Good for

About Wan 3.0

What it is

Wan is Alibaba's video generation family. It has two tracks. Earlier versions such as Wan 2.1 and 2.2 were published as open weights under Apache-2.0, and they became the default open video models in ComfyUI. The newest version, Wan 3.0, entered beta on Alibaba Cloud Model Studio in August 2026 and is offered as an API. We found no open weights for Wan 3.0.

Key features

  • Wan 3.0: up to 30 seconds per clip, double the 15 seconds of Wan 2.7
  • Accepts text, images, video, audio and even documents (PDF, PPT) as input
  • Multilingual voices with synchronised facial expressions
  • Instruction-based editing of existing clips
  • Wan 2.2 (open): 480p/720p text-to-video, image-to-video, speech-to-video and character animation

Where you can use it

Wan 3.0 is available on Alibaba Cloud Model Studio and Qwen Cloud, and in ComfyUI partner nodes (Wan 3.0 and 2.7). Wan 2.6/2.7 is documented in Higgsfield, Krea, Leonardo.Ai and invideo AI. The open Wan 2.x weights run locally in ComfyUI.

Pricing and rights

Wan 2.1/2.2 weights are free to use under Apache-2.0, including commercial use. You need your own GPU. Alibaba had not published Wan 3.0 API pricing at the beta announcement. In partner apps, Wan uses that app's credits.

Who it is for

Tinkerers and studios that want a capable open model to run and fine-tune locally. Also developers who want a long-clip API from a major cloud.

Verdict

Wan is the most practical open video family, and the Apache licence is creator-friendly. The best new features (Wan 3.0, editing, multi-reference) are API-only. Details like resolution and pricing for 3.0 were still thin in beta.

open-weights text-to-video image-to-video comfyui alibaba

Pros

  • Apache-2.0 open weights for Wan 2.1/2.2
  • Wan 3.0 makes 30-second clips with voices
  • Huge ComfyUI community (LoRAs, workflows)
  • Documents and web pages accepted as input

Cons

  • Wan 3.0 is API-only (no open weights found)
  • Wan 3.0 pricing and resolution not published at beta
  • Local runs need a strong GPU

Similar AI models

All video models →

Veo 3.1 🧠 ModelPaid

Google DeepMind's text- and image-to-video model with native sound, up to 4K

Google DeepMind · 2025-10

8.8 Visit ↗

Kling 3.0 🧠 ModelFreemium

Kuaishou's video model with native multilingual audio, lip sync and 15-second clips

Kuaishou · 2026-02

8.6 Visit ↗

Seedance 2.5 🧠 ModelPaid

ByteDance's video model: 30-second clips with audio, many references and local edits

ByteDance Seed · 2026-07

8.6 Visit ↗

Runway Gen-4.5 🧠 ModelFreemium

Runway's own video model, now with native audio and one-minute multi-shot scenes

Runway · 2025-12

8.2 Visit ↗

LTX-2.5 🧠 ModelOpen source

Lightricks' open-weight video model with synchronised audio and 4K output

Lightricks · 2026-08 · open weights

7.8 Visit ↗

Popular searches