AI Models in ComfyUI
The AI models you can choose inside ComfyUI, grouped by type. Each model page shows the other apps that offer it, so you can compare prices for the same model.
Updated September 2026 · 24 listings
How to pick a model in ComfyUI
Pick the model from the app's model selector (often under the prompt box or in settings). Credit costs usually differ per model, and some apps only offer the newest models on paid plans.
Claude Fable 5.1 🧠 ModelFreemium
Anthropic's Claude models (Fable 5.1, Opus 5.5, Sonnet 5, Haiku 4.5) with 1M-token context
GPT Image 2.5 🧠 ModelFreemium
OpenAI's image model behind ChatGPT Images, with 4K output and up to 16 reference images
GPT-6 Astra 🧠 ModelFreemium
OpenAI's GPT-6 family (Astra, Sol, Luna) behind ChatGPT and the OpenAI API
Nano Banana 2 🧠 ModelFreemium
Google's Gemini image models (Nano Banana 2, Pro and 2 Lite) for generation and editing up to 4K
Eleven v3 🧠 ModelFreemium
ElevenLabs' expressive text-to-speech model with audio tags and 70+ languages
Veo 3.1 🧠 ModelPaid
Google DeepMind's text- and image-to-video model with native sound, up to 4K
Gemini 3.1 Pro 🧠 ModelFreemium
Google's Gemini models (3.1 Pro, 3.8 Flash) in the Gemini app, NotebookLM and the Gemini API
FLUX.2 🧠 ModelOpen source
Black Forest Labs' FLUX.2 family: hosted [max]/[pro]/[flex] plus open-weight [dev] and [klein]
Kling 3.0 🧠 ModelFreemium
Kuaishou's video model with native multilingual audio, lip sync and 15-second clips
Seedance 2.5 🧠 ModelPaid
ByteDance's video model: 30-second clips with audio, many references and local edits
Seedream 5.0 Pro 🧠 ModelFreemium
ByteDance's Seedream image models, used in CapCut, Dreamina and many creative apps
MiniMax H3 (Hailuo 3.0) 🧠 ModelFreemium
MiniMax's open-weight video model: native 2K, stereo sound, 15-second clips
Runway Gen-4.5 🧠 ModelFreemium
Runway's own video model, now with native audio and one-minute multi-shot scenes
Wan 3.0 🧠 ModelOpen source
Alibaba's video family: open-weight Wan 2.x and the 30-second Wan 3.0 API
Ideogram 4.0 🧠 ModelOpen source
Ideogram's text-rendering image model with JSON prompting, now also as non-commercial open weights
LTX-2.5 🧠 ModelOpen source
Lightricks' open-weight video model with synchronised audio and 4K output
Qwen-Image 3.0 🧠 ModelOpen source
Alibaba's Qwen image models, known for dense multilingual text rendering, with open-weight versions
Luma Ray3.2 🧠 ModelPaid
Luma's HDR video model with up to 16 keyframes in a 20-second 1080p clip
Grok 4.7 🧠 ModelFreemium
xAI's Grok models, built into Grok and X, with a 500K-token context on Grok 4.7
ACE-Step 1.5 🧠 ModelOpen source
MIT-licensed open music model that makes full songs with vocals locally in seconds
Vidu Q3 🧠 ModelFreemium
ShengShu's video model with native audio and 16-second multi-shot scenes
PixVerse V6 🧠 ModelFreemium
PixVerse's video model: 1-15 s clips at 1080p with native audio and multi-shot
HunyuanVideo 1.5 🧠 ModelOpen source
Tencent's 8.3B open-weight video model that runs on a single consumer GPU
Stable Diffusion 3.5 🧠 ModelOpen source
Stability AI's open-weight image models (SD 3.5 Large, Turbo, Medium) for local and custom workflows