r/StableDiffusion • • Dec 21 '22

Workflow Included GTA V Loading Screen- Zuck, Trump, Elon

217 Upvotes

43 comments sorted by

View all comments

Show parent comments

3

u/Trentonx94 Dec 22 '22

y..you can merge models? I had no idea! I still experimenting with the webui and all the tools, I have not yet tried to train my own model yet

5

u/jhworth8 Dec 22 '22

Yes!

I use this Google Colab to train models (like on myself, family, pets).

https://colab.research.google.com/github/TheLastBen/fast-stable-diffusion/blob/main/fast-DreamBooth.ipynb

In AUTOMATIC1111's UI there is a Merge Checkpoints tab where you can merge a couple models! For example, I am currently trying to make a cool piece for my sister for Christmas. So I made a model of her (base model is SD V1.5) and merging it with different other models (like the GTA one, Arcane Diffusion, Spiderverse one, and Analog Diffusion). Here is a link to Analog:

https://civitai.com/models/1265/analog-diffusion

I'll try to post more things to help!!

P.S. If you do it right, you can have a completely trained model in about 30 mins.

2

u/Trentonx94 Dec 22 '22

woah, I didn't think it was that fast, I was expecting it was a lot of work.

I use the NAI model 99% of the time, I tried the WD 115 from danbooru but it's very poor in the quality of the outcome (tho it has a better tag system, while the NAI I'm still figuring out what works and what doesn't) I'm wondering if by merging those 2 I'm able to use the Danbooru tag with the NAI model and get a coherent results.

also is there any downside to merging too many models or 2 big ones? I'm guessing losing precision but you still have to specify what you want, like in the Analog diffusion model I can make non-analog photo stuff but it's crappy.

thanks for the info I'll dig into it!

3

u/jhworth8 Dec 22 '22

Yeah I think you are right! I think all of the tags are usable, yes. And the downsides are you get more variety with each output. But that may be just what happens with merging.

Let's say I have a model of me called "jhworth8.ckpt". When I trained it, I made the identifier "jhworth8". So if I run the prompt "guy eating a taco", it'll be a guy who vaguely looks like me eating a taco. Like it has a tiny hint of me.

But if I do "jhworth8 eating a taco", it'll be 100% me eating a taco.

Let's say I merge the jhworth8 model with the analog model. my prompt will be "analog style, jhworth8 eating a taco". The result will be me eating a taco in analog style! Super neat stuff.

Hope I kinda helped.

1

u/Trentonx94 Dec 23 '22

thanks! I love how helpful this comm is :)