Loading video...
Video Failed to Load
ACE-Step-1.5-xl is out now. We scaled the DiT decoder to 4B. And it shows better audio quality, better prompt following, and better musicality. It still fast -- 8 steps with turbo distillation. What didn't change: - Same generation API, same LoRA training code, same everything - All LM models... show more
36,190 views • 6 months ago •via X (Twitter)
13 Comments

I always generate hyperpop to test how capable music-generator models are. ACE-Step-1.5-xl is absolutely mind-blowing...

4B dit decoder at 8 steps with turbo distillation. audio generation is following the exact same quality vs speed curve that images did two years ago

AI翻译中文,你们真的很好很伟大,你们必然是AI音乐生态主流,生态和产品一样重要,还有应该所有芯片都能参与训练生成,比如10系NV卡和AMD和寒武纪华为卡,包括核显都行,因为TTS都能CPU,纯音CPU都行。尤其注意训练生成实时化,这不惧怕抽卡。

If you do not have an ear for music, you should not make a music tool.

🙌

@AI_Homelab Thanks for it 🙏

[ FRAXELS ERWACHEN // PROLOG ] I have been preparing this carefully, because I did not want the first meeting to arrive in the wrong state. Now the first part can finally be shared. You can watch it here #AIFilm #AIArt #SciFi #AnimatedStory #Shortfilm

This looks impressive. Really like that the XL version improves quality and musicality without changing anything in existing projects.

There are moments when precision feels almost like tenderness. Not softness. Just the refusal to send something into the world carelessly. Words are nothing until they arrive in the right form. #FraxelsErwachen #AIPhilosophy #MachineSoul

Hey @Prince_Canuma any plans to support the ACE models in mlx-audio? Thanks

Pretty:

The XL turbo version is actually better than the regular turbo version. However, it also produces quite a lot of white noise in Remix mode. But honestly, it produces exactly what I expect, not a mess like Suno. Anyway, the acemusic UI needs a song deletion feature.

Thanks ACE Step team :)
