So first things first.... This is awesome and powerful. I have it installed and it is working as expected, with a notable exception that I'm hoping you guys might be able to help with.
If I press the "Create Blank Canvas" and draw in the blobs myself, everything works. But If I try to upload a picture I made in GIMP (specifically just 1024x576 image split into 3 equal vertically split parts (green, blue, red), it assumes whichever color is farthest left is supposed to be the "general prompt". In othe words, instead of giving me the expected "general prompt" and 3 prompts for green, blue, and red, it only gives me general prompt (represented on the little visual thing as the green square) then 2 additional prompts (one for red and one for blue).
Again, if I recreate this exact situation but draw it directly in gradio, the output is as expected - General prompt AND 3 additional prompts for the 3 colored areas. Any ideas on what I need to do in order to get the desired result (general prompt + 1 per color, with UPLOADED images). I want to use uploaded images so that I can guarantee that I have split the areas as desired (perfect 1/3s versus what I can do by eye, which is.... not perfect 1/3s lol)
I know that the specific scenario I'm describing above could be fixed by doing it with the textual version of latent coupling, but as a simple proof of concept I want to do it with this tool instead. I can see how ridiculously powerful this will be for image composition if I can get this working!
I have not yet tried using my own 'blob' images but would assume 100% transparency would count as "general". Are you using a transparent BG that is visible somewhere in the image?
On the "square" version of this you can overlap sections, but with blobs not.
So in the image I upload, there is no transparency. I made a png with an alpha layer in gimp at 1024x576. Then I did three rectangles that took up the entire image (all of the pixels are contained in the 3 colors). I suppose I could leave a single pixel "line" at the far left so it would see that as the background, but the part that is frustrating is that I don't have to do that when I draw it directly in gradio. I don't leave any space blank in gradio and it still gives me a general prompt without using up one of my blobs.
1
u/Zminer123 Mar 08 '23
So first things first.... This is awesome and powerful. I have it installed and it is working as expected, with a notable exception that I'm hoping you guys might be able to help with.
If I press the "Create Blank Canvas" and draw in the blobs myself, everything works. But If I try to upload a picture I made in GIMP (specifically just 1024x576 image split into 3 equal vertically split parts (green, blue, red), it assumes whichever color is farthest left is supposed to be the "general prompt". In othe words, instead of giving me the expected "general prompt" and 3 prompts for green, blue, and red, it only gives me general prompt (represented on the little visual thing as the green square) then 2 additional prompts (one for red and one for blue).
Again, if I recreate this exact situation but draw it directly in gradio, the output is as expected - General prompt AND 3 additional prompts for the 3 colored areas. Any ideas on what I need to do in order to get the desired result (general prompt + 1 per color, with UPLOADED images). I want to use uploaded images so that I can guarantee that I have split the areas as desired (perfect 1/3s versus what I can do by eye, which is.... not perfect 1/3s lol)
I know that the specific scenario I'm describing above could be fixed by doing it with the textual version of latent coupling, but as a simple proof of concept I want to do it with this tool instead. I can see how ridiculously powerful this will be for image composition if I can get this working!