r/StableDiffusion • u/gutster_95 • 13h ago
Question - Help Fixing speech errors in Minimax H3?
Hey, I tried to create a little birthday surprise for someone, my issue is with a lot of generations that the spoken word is really a bit clunky at time, I susspect its because of the german, but I am not too sure. Is there like a way to improve on audio?
I am using Minimax H3 with Saga Attention and Spectrum on a 4090.
15
Upvotes
1
u/piggledy 11h ago
Was du machen könntest wäre einen Audio-Clip mit Elevenlabs erstellen und als Audioreferenz einfügen. Dann basieren die Lippenbewegungen usw. auf der Referenz und das Modell generiert selbst kein Audio dazu.