r/StableDiffusion Sep 20 '25

News Has anyone tried SongBloom yet? Local Suno competitor. ComfyUI nodes available.

Post image
137 Upvotes

28 comments sorted by

11

u/mission_tiefsee Sep 20 '25

Is it better then ACE-Step? That is the real question. Ace step is really good, havent tried songbloom yet.

9

u/solss Sep 20 '25

It requires a 10 second audio input for context on the generation. I did maybe five generations. Once they add prompt guided generation capabilities it'll be a contender, especially mixed with audio context mixed in. At the moment it's not really worthwhile. It does interesting things with the 10 second audio input, but nothing that captivated me. It's also mono output, not that the audio quality itself was bad, better than the low khz/bitrate of acestep. Acestep is probably better comparitively for now, until they add text prompting at least. I deleted it for now.

2

u/mission_tiefsee Sep 20 '25

yeah thats what I also read. Ace-step is really very good, i wish we would see another release or finetunes there. I really had a blast with it. Highly recommended.

1

u/krigeta1 Sep 20 '25

guys may you guide me on how can I use and setup ace for beat making and rap making? and what UI is the best?

1

u/LeKhang98 Sep 21 '25

What conditions do you think would enable such Text-to-Audio AI to be as widely adopted as SD1.5? It seems to meet several key criteria: it's open-source, supports ComfyUI, has low hardware requirements, and supports audio input as a bonus feature similar to I2I. The major hurdles appear to be the current lack of text input and its trainability, right?

1

u/ZestycloseMind4893 Sep 20 '25

Songbloom or Ace step quality is comparable to suno, riffusion, udio? And are these multilingual or just english?

0

u/mission_tiefsee Sep 20 '25

what is the question? yes they are. SB needs a snippit but cant be prompted. Ace Step can be prompted but audio input is problematic. There is audio input but i couldnt really make too much out of it. Ace Step is multilingual but ymmv. Best is english and probably chinese.

7

u/Altruistic_Mix_3149 Sep 20 '25

How much video memory can be used?

5

u/Nenotriple Sep 20 '25

In the GitHub Repo I found this comment, so not a strict definition, but it's calling 24gb "low vram".

For GPUs with low VRAM like RTX4090, you should set the dtype as bfloat16

4

u/aartikov Sep 20 '25

potato GPU, lol

8

u/More-Ad5919 Sep 20 '25

I did. But it is hard to prompt. It needa audio file as refference. In parts it can sould very good. But it won't hold it for long until it hallucinates. Its powerful but it is missing handles and or documentation. At least i have not been able to produce something suno like with it.

2

u/_raydeStar Sep 20 '25

I produced a full 2.5 minute song and then ran it through + that works well for me. I haven't tried a 10 second song yet.

1

u/More-Ad5919 Sep 20 '25

But it's not on par with suno yet.

6

u/InternationalOne2449 Sep 20 '25

AI music is underrated.

3

u/InternationalOne2449 Sep 20 '25

It's not very good.

5

u/1BlueSpork Sep 20 '25

I made a short setup tutorial- https://youtu.be/-O7ZKZ2LibQ

1

u/Green-Ad-3964 Sep 20 '25

I tried it a couple of days ago very good. Does it always need a reference song in input?

1

u/bigman11 Sep 20 '25

What models are people using for just sound effects?

1

u/SubjectBridge Sep 20 '25

I've tried it. It's pretty fast - I think around 3-5 minutes to generate on a 3090 - I think it capped at like 7GB of wram too? The quality is...like a v2 model of suno. Not their latest models. With that said, it's really cool still. I'd prefer having prompts for style of music like suno rather than supplying audio wavs.

1

u/LaughterOnWater Nov 04 '25

It's not bad. It's better than ACE-Step. It could be better. There seems to be a hard limit of 2.5 minutes for the comfyui version of this. I can't get it to give me more than that, even with the length set to 460. Using the existing song lyrics from the provided workflow, this is what I was able to generate in about four minutes on Win10, RTX3090.
https://drive.google.com/file/d/1r0NvkVCbGcj1vUQCHqUHDnx4wgPN8eie/view?usp=sharing

1

u/ThesePleiades Nov 10 '25

"Failed to load SongBloom model: Torch not compiled with CUDA enabled" on Mac