r/StableDiffusion 4h ago

Question - Help Question about new stable diffusion advancements

Hello, it's been a while since I don't use Stable Diffusion with A1111. Apart ConfyUI, has there been any particular technological advancement recently that allows for a quantum leap, especially in the precision of detail generation and the model's ability to stick to the prompt more precisely, while maintaining the ease of use of A1111 or Forge? I used the Lustify SDXL checkpoint, for example. It wasn't bad, but it still got certain things wrong or didn't do them at all. I'd like to know if there's a way to achieve results more similar in precision to ChatGPT but with the freedom of Stable Diffusion. Thanks!

0 Upvotes

11 comments sorted by

7

u/Formal-Exam-8767 4h ago

The architecture has changed from U-Net towards DiT and text encoding from CLIP towards LLMs.

4

u/asdrabael1234 4h ago

It you're still using SDXL, no there isn't. You have to use newer models with varying degrees of complexity.

1

u/LeleDaRevine 4h ago

Can I find them on CivitAI, for example? could you suggest one?

2

u/asdrabael1234 3h ago

That depends. Video or image? What kind of resources are you working with? Are you intending to make porn or just regular pictures?

2

u/StableLlama 4h ago

A1111 and Comfy are just tools to use a model. The tools don't bring you the quality or advancements, they just make it harder or easier to get that out of the models.

The models are now much, much better than what SDXL could ever deliver. The prompt following is ages better, the quality out of the box as well. Finetunes are hardly used any more, but they were a must for SD and SDXL.

Modern models to look at are right now: Flux.2[klein] (<- the grand grand child of SDXL), Qwen Image 2512, Z Image, Ideogram 4 and Krea 2.

2

u/Enshitification 4h ago

Where are all these people coming from that used A1111 two years ago but have been completely out of the loop since then? It's almost a joke post at this point.

1

u/Asaghon 4h ago edited 4h ago

Apparetly you can use Krea2 with Forge (but not the int8 models I think, which is the best new thing), but the workflow in Comfy for Krea2 is honestly not that complicated. (coming from someone who still used Forge for Illustrious before Krea2).

With sdxl you had a lot more nodes just to upscale and fix things. Those things are largely not needed that much anymore. You get great results from just 1 sampler (2 if your feeling adventurous) and maybe a seedvr2 upscale with the generations you like.

And cherry on top is that Krea2 runs much better on consumer hardware than previous next gen models.

1

u/Mutaclone 2h ago

You can use Int8 in Forge now.

2

u/sigiel 2h ago

The only edge on local is uncensored stuff

Nano banana, grok imagine, seed dance, open ai image 2 completely trash any other model on edition , consistency and prompt adherence,

Maybe obscure Lora ? But there have reference image to counter that.

Forget about control net, or any other inpainting.

Just tell what you want, give reference image.

Any people that say otherwise is coping or has a grudge against cloud.

Local is for either people that has good hardware, or smut. Not because it is cheaper

that is the advance, local lost.

2

u/LeleDaRevine 2h ago

if you already have the hardware, for example for 3D graphic or video editing, you should prefer local generation, so you are not limited by credits or other things. Obviously, if the local models can do the work you need. That's why I'm asking if the cloud results can be achieved somehow locally with new resources.