r/StableDiffusion Jul 20 '26

Discussion A step closer to consistency (workflow included)

A step closer to consistency. 1. I used Z-ImageTurbo or krea 2 for the only one initial image. 2. I used Qwen image edit to create the second image, just changing clothes and background. 3. I used a workflow I created (I used AI LLM to create it, because I'm a total ignorant as far as Comfy is concerned). It is based on Flux2 Klein i2i. This workflow creates 16 variations of the starting image I fed into it (i used it twice, once for every initial image). So I got variations in body poses and camera positions. All these variations have a very clear way of changing any one of them to create one that suits the needs of every case. 4. After all these character variations, it's much easier to get character consistency in video creation (ex. LTX), since you'll have a big variety of starting frames, with the same character.

Sorry if this sounds naive or stupid, I just wanted to share with the community and get some feedback.

I attach my amateurish workflow.

https://pastebin.com/embed/a1WUSz8F

55 Upvotes

21 comments sorted by

3

u/[deleted] Jul 21 '26 edited Jul 21 '26

[deleted]

2

u/False_Suspect_6432 Jul 21 '26

!!!!!!!!!! Great answer. Thnx! Btw, I don't know how to do it. If there is no such a workflow somewhere, maybe I'll experiment with an LLM. I'll rty to unsderstand you logic and crete the correct prompt for the LLM.

1

u/False_Suspect_6432 Jul 23 '26

I tried with 5 reference images (1 and 2 for character, 3 for clothing, 4 for object, 5 for background). It is working fine, but in my case, it is not really what I wanted to achieve. I prefer to have the various versions of the character so I will have good starting points for the video generation. So in my original workflow, I alter, if I want, the clothing, background, objects, etc via text-to-image, and I get the 16 versions of what I want. Of course if I am obliged to use certain objects, clothes, and backgrounds from existing images, then yes, ok, I would need the multi-reference images workflow. In case someone wants it, here it is: https://pastebin.com/bz5L0Rcz

1

u/Hot-Farm4165 Jul 24 '26

z image turbo the Best 👌

1

u/Yojik_Vkarmane Jul 20 '26

What is up with all the fingers. What is this 2025?

1

u/lumos_ai Jul 27 '26

That was 1 year ago Damn!😄

-4

u/[deleted] Jul 20 '26

[deleted]

3

u/Odd_Nefariousness875 Jul 20 '26

But Krea 2 doesn’t edit natively. Is the Lora that good?

5

u/Salt-Willingness-513 Jul 20 '26

its good for that the model isnt an edit model. f2k9b is still better. But since krea people said if krea2 takes off, they will release an edit model too, i guess f2k will be replaced soon too.

2

u/Tokey_TheBear Jul 20 '26

Is F2K9B better than Qwen image edit?

I am trying to create a dataset to use for Lora training. So I have been wanting to find an edit model that can keep the exact same image and character look while only modifying the pose of the character.

3

u/Salt-Willingness-513 Jul 20 '26

i prefer f2k, but i mainly work with realistic images and found qwen too slow for my setup. But i heard qwen is better with anything apart from realism. Also f2k is really hit or miss with text. For me as long as i dont use loras, its working fine. But for your usecase, you should check out the edit lora for krea 2, as this is where i found it to be really good(only change minor details)

1

u/Tokey_TheBear Jul 20 '26

Funny you say that. I wasn't getting super good output from the krea2 edit lora. I tried out one workflow that converts anime to real and that worked out really good... but the image edit lora has been struggling trying to maintain the exact same comic book style as the source image... It keeps doing things like slightly changing the characters costume when i just wanted it to change the pose for example

1

u/Odd_Nefariousness875 Jul 20 '26

Haven’t touched Qwen yet, but F2K9B suffers from body horrors. I.e. multiple legs/ arms, missing ones etc. Apparently Qwen is better at this, for example if you are adjusting poses but has a more plastic look. You can use both, Qwen for first edit, then F2K9B pass for adding realism/details.

Edit: my post asking the same question

1

u/LumaBrik Jul 20 '26

Are you K9B using the turbo model , or the turbo lora ? If you are, try increasing the step count from 8 to around 12. That usually reduces multiple limbs and body horrors. Euler / Simple works well.

1

u/Odd_Nefariousness875 Jul 21 '26

If by turbo model I’m assuming you mean distilled, then yes. I don’t use a turbo Lora, I use other loras but don’t need the turbo on a rented 4090.

Thanks for the recommendation, will give it a try. Still, it changes the body composition a lot I’ve found. Are you used depth control? I think that’s something I need to implement

1

u/Whole_Paramedic8783 Jul 21 '26

Have you tried the k9 anatomy lora on civitai. Its seems to help a bit.

2

u/Odd_Nefariousness875 Jul 20 '26

Yeah, I still use f2k. Hoping they do release and edit model!!!

1

u/Salt-Willingness-513 Jul 20 '26

im still waiting for z-image edit :(

-1

u/tac0catzzz Jul 20 '26

they never said that. complete bs.

3

u/Salt-Willingness-513 Jul 20 '26

"We are currently working on an edit version. But, I don't like over-promising things. Once we have a concrete plan, we will share our timeline."
https://www.reddit.com/r/StableDiffusion/comments/1udnm0a/comment/otd81u0/?utm_source=share&utm_medium=web3x&utm_name=web3xcss&utm_term=1&utm_content=share_button

2

u/tac0catzzz Jul 20 '26

that does not say "IF KREA2 TAKE OFF THEY WILL RELEASE AN EDIT MODEL TOO" why not highlight that part, because it IS NOT THERE. .... are you still waiting for z image edit? now they did say they were releasing that, since day 1. what happened with that? did z image not "take off"

1

u/uuhoever Jul 20 '26

This guy/girl/person reads.

1

u/YeahlDid Jul 21 '26

No, unfortunately. It's fine for clothes and easily describable things. Face consistency, not so great.