Try “Veo”, “voiceover”, “thumbnail” or “Descript” · Esc to close

ACE-Step 1.5 🧠 Model Open source

MIT-licensed open music model that makes full songs with vocals locally in seconds

Music & audio models · Open source ★ 13k · MIT · updated 2026-09-03

7.4editor score
View on GitHub ↗
GitHub stars
13k
Stars this week
–
Forks
1.7k
Licence
MIT
Last push
2026-09-03
Maintainer
ACE-Step

Where you can use ACE-Step 1.5

Apps in this directory that let you generate with ACE-Step 1.5, from their own documentation.

See every app × model on the model map →

Good for

About ACE-Step 1.5

What it is

ACE-Step is an open-source music generation model. ACE-Step 1.5, released on 2 April 2026, generates songs from 10 seconds to 10 minutes long, with vocals and lyrics in 50+ languages. It is published under the MIT licence. An XL variant with a 4B-parameter decoder improves quality for users with bigger GPUs.

Key features

  • Songs from 10 seconds to 10 minutes, with vocals and lyrics
  • 50+ lyric languages
  • A full song in under 2 seconds on an A100, or under 10 seconds on an RTX 3090
  • Runs in 6 GB or less of VRAM with quantisation and CPU offload. 12 GB+ is recommended for XL
  • Official ComfyUI integration and a community ecosystem

Where you can use it

Clone the GitHub repo and run it locally with uv, or use the official ComfyUI nodes. We found no commercial creator app in our directory that documents ACE-Step by name.

Pricing and rights

Free under the MIT licence, including commercial use of the model. You pay only for your own GPU. The licence covers the software. You are still responsible for not imitating real artists or reproducing copyrighted lyrics in what you publish.

Who it is for

Musicians, ComfyUI users and developers who want an open, commercial-friendly song generator with vocals on their own hardware.

Verdict

ACE-Step 1.5 is the strongest permissively licensed song model with vocals, and it is fast. Vocal realism and mixing still trail Suno v6 and Lyria 3.5, and setup takes some technical comfort.

open-source text-to-music vocals comfyui local

Pros

  • MIT licence
  • Vocals and lyrics in 50+ languages
  • Very fast generation
  • Runs on modest GPUs

Cons

  • Vocals behind Suno/Lyria
  • Technical setup
  • No hosted consumer app from the makers

Suno v6 🧠 ModelFreemium

Suno's song model: full tracks with vocals and lyrics from a prompt

Suno · 2026-09

8.6 Visit ↗

Lyria 3.5 🧠 ModelPaid

Google DeepMind's music model for full songs with vocals, lyrics and SynthID

Google DeepMind · 2026-07

8.0 Visit ↗

Eleven Music v2.5 🧠 ModelFreemium

ElevenLabs' music model: songs up to 5 minutes with vocals, cleared for most uses

ElevenLabs · 2026-09

7.9 Visit ↗

Stable Audio 3.0 🧠 ModelOpen source

Stability AI's licensed-data audio models up to six minutes, with open-weight tiers

Stability AI · 2026-05 · open weights

7.5 Visit ↗

MusicGen (AudioCraft) 🧠 ModelOpen source

Meta's open research model for short instrumental music from text or a melody

Meta (facebookresearch) · 2023-06 · open weights

Popular searches