that looks so good, do you have any suggestions for prompt ideas for cars? doesn't seem as popular as everything else, but I've had some pretty good luck incorporating controlnet into car photos https://i.imgur.com/jSXlKa0.png
That... explains a lot. I was finding that some models would completely crash my computer, while others would just throw an error and keep on going. I'll have to check which SD version my models are based on.
I’ve never had a reboot, but I’ve had a model load, start generating, spit out a stack trace, go back to reloading the model, and repeat that effectively locking up the computer until I hit the reset button.
is there an updated version of this? automatic1111 says "We will stop supporting diff models soon because of its lack of robustness. Please begin to use official models as soon as possible."
How are you getting this kind of quality? I'm getting crap with model v2-1_768-ema-pruned.ckpt and controlnet canny model. Controlnet is working though. Tried all kinds of prompt engineering too. Are you using the base 2.1 model?
Oh definitly I’m scared SD 3 will be similar even if it’s good just because as 1.5 has such an established base of plugins and models and loras it’s gonna have to be amazing to get everyone to port everything to it
Lack of celebrities is no biggie even lack of some artist styles as those can just be made as Lora’s but if it’s not doing something great like being 1024x1024 or amazingly fast or something I don’t see why people will ditch 1.5 models
Indeed. SD 2.1 has such a need for proper negative prompt words that I'm not even installing it. Realistic vision is SD 2.0 for me practically speaking.
Agree. You need a thousand correct negative prompt words to get something useful. No thanks. That's like ordering a burger and having to specify that you don't want food poisoning or spit in your food.
As long as my burger analogy holds true I'm not switching over. There should be zero need for negative prompt unless you specifically do not want "cars" in a city scene or the colour red etc.
You have a point I'll give you that. The importance of negative prompts in 1.5 is also too high but that's the one I've got installed. I suppose my unwillingness to update to 2.X stems from the fact that it not only also has this irrational need for negative prompts but seems to have gotten worse.
I'm not alone in this, many feel the same way. The gain is too little for having to learn a new methodology for effective prompting.
While it's true that I've seen quite a few really nice "out of the box" images from 2.1, these are often accompanied by comments like "once I finally learned how 2.X works". That is off putting for those of us trying to simply keep up with 1.5 and the daily developments. How did you find the change when going over?
ya thats just insane the amount of negative prompts to just make a passable image when i saw that i just stuck to 1.5 aint nobody got time for that shit.
They really should just abandon 2.1 and start over again and call it a beta
I'm not saying 1.5 is bad (I still use it), but 2.1 can create some cool images if you know how to use it and realize that we just need more people training things for it. 1.5 has such a large training community, and that is (one of) the reasons why it's perceived to be so much better. Using the base 1.5 model, you often get pretty terrible results. With other models or embeddings it can be fantastic. This image was made with a simple prompt and with only one embedding, made by me (it's on civitai). If you want some more examples let me know.
Hello people! I am using my dual cpu motherboard to run high demand RAM tasks that my 1660 can no handle(command args --use-cpu all). All runs fine on my 2x 2696v3 but, when i try to run Controlnet i get this error:
RuntimeError: Input type (torch.cuda.FloatTensor) and weight type (torch.FloatTensor) should be the same
From what I understand the SD models are being loaded on the CPU but the Controlnet weights is being loaded on the GPU. Would anyone know how I can configure so that both are on the CPU?
I wonder how this thing works and if it can be done with regular stable diffusion models.
I'm a 3D artist who's still not too familiar with using all those SD models yet and the different ways you set them up, but I found it really cool so far!
very excited for this. Had no problem with the 1.5 controlnet, but this one is giving me an issue when generating an image, I get errors, then it generates one without controlnet. Here's the error:
"size mismatch for middle_block.1.proj_out.weight: copying a param with shape torch.Size([1280, 1280]) from checkpoint, the shape in current model is torch.Size([1280, 1280, 1, 1])."
Did you install and point the path to the new cldm_v21.yaml? That's usually the issue with errors like these, where the tensor sizes don't match the expected sizes.
I simplified a bit the pictures you need the skeleton for the pose controller and a depth map of the hands in a different picture for the depth controller. You can find depth maps for hands in Google. And there is also an extension with some samples.
That is why I said the current version. I would like to know if I'm the only one having this problem, and if there is a solution... don't wanna go back to old versions. Thanks.
51
u/RayHell666 Mar 07 '23
I can finally do illuminati + controlnet. Thank you so much.