ACE-Step 1.5 🧠 Model Open source
MIT-licensed open music model that makes full songs with vocals locally in seconds
- GitHub stars
- 13k
- Stars this week
- –
- Forks
- 1.7k
- Licence
- MIT
- Last push
- 2026-09-03
- Maintainer
- ACE-Step
Where you can use ACE-Step 1.5
Apps in this directory that let you generate with ACE-Step 1.5, from their own documentation.
Good for
About ACE-Step 1.5
What it is
ACE-Step is an open-source music generation model. ACE-Step 1.5, released on 2 April 2026, generates songs from 10 seconds to 10 minutes long, with vocals and lyrics in 50+ languages. It is published under the MIT licence. An XL variant with a 4B-parameter decoder improves quality for users with bigger GPUs.
Key features
- Songs from 10 seconds to 10 minutes, with vocals and lyrics
- 50+ lyric languages
- A full song in under 2 seconds on an A100, or under 10 seconds on an RTX 3090
- Runs in 6 GB or less of VRAM with quantisation and CPU offload. 12 GB+ is recommended for XL
- Official ComfyUI integration and a community ecosystem
Where you can use it
Clone the GitHub repo and run it locally with uv, or use the official ComfyUI nodes. We found no commercial creator app in our directory that documents ACE-Step by name.
Pricing and rights
Free under the MIT licence, including commercial use of the model. You pay only for your own GPU. The licence covers the software. You are still responsible for not imitating real artists or reproducing copyrighted lyrics in what you publish.
Who it is for
Musicians, ComfyUI users and developers who want an open, commercial-friendly song generator with vocals on their own hardware.
Verdict
ACE-Step 1.5 is the strongest permissively licensed song model with vocals, and it is fast. Vocal realism and mixing still trail Suno v6 and Lyria 3.5, and setup takes some technical comfort.
Pros
- MIT licence
- Vocals and lyrics in 50+ languages
- Very fast generation
- Runs on modest GPUs
Cons
- Vocals behind Suno/Lyria
- Technical setup
- No hosted consumer app from the makers
Similar AI models
All music & audio models →Suno v6 🧠 ModelFreemium
Suno's song model: full tracks with vocals and lyrics from a prompt
Lyria 3.5 🧠 ModelPaid
Google DeepMind's music model for full songs with vocals, lyrics and SynthID
Eleven Music v2.5 🧠 ModelFreemium
ElevenLabs' music model: songs up to 5 minutes with vocals, cleared for most uses
Stable Audio 3.0 🧠 ModelOpen source
Stability AI's licensed-data audio models up to six minutes, with open-weight tiers
MusicGen (AudioCraft) 🧠 ModelOpen source
Meta's open research model for short instrumental music from text or a melody