Try “Veo”, “voiceover”, “thumbnail” or “Descript” · Esc to close

AI Models for Content (2026): 5 Video, Image, Voice & Text Models and Where to Use Them

The generative models behind the apps: language, image, video, voice and music models, and the apps where you can use each one.

The same model is often sold by several apps at different prices and limits. Each model page lists where you can use it; the model map compares them all.

Veo 3.1 🧠 ModelPaid

Google DeepMind's text- and image-to-video model with native sound, up to 4K

Google DeepMind · 2025-10

8.8 Visit ↗

Kling 3.0 🧠 ModelFreemium

Kuaishou's video model with native multilingual audio, lip sync and 15-second clips

Kuaishou · 2026-02

8.6 Visit ↗

Seedance 2.5 🧠 ModelPaid

ByteDance's video model: 30-second clips with audio, many references and local edits

ByteDance Seed · 2026-07

8.6 Visit ↗

Runway Gen-4.5 🧠 ModelFreemium

Runway's own video model, now with native audio and one-minute multi-shot scenes

Runway · 2025-12

8.2 Visit ↗

Wan 3.0 🧠 ModelOpen source

Alibaba's video family: open-weight Wan 2.x and the 30-second Wan 3.0 API

Alibaba (Tongyi Wan) · 2026-08 · open weights

8.0 Visit ↗

Popular searches