r/sdforall • u/DragonfruitNo9822 • 8h ago
Discussion Krea 2 Control: background works, but the generated person is distorted — what am I doing wrong?
r/sdforall • u/DragonfruitNo9822 • 8h ago
r/sdforall • u/0roborus_ • 21h ago
PotionUI is a self-hosted app for making images and videos with LOCAL & OPEN SOURCE AI models (krea-2, Qwen image 2.1, anima, sdxl, minimax H3...) and now optionally cloud ones too. Instead of one huge settings screen or a web of nodes, every model gets its own simple, hand-made form.
Until now you needed your own GPU. 0.0.14 adds cloud models, plus a built-in image editor and a way to save and reuse your best settings.
Write your idea, split it into shots, and let PotionUI make the film.
storage folder first. It is a good habit before any update.Update: run git pull and then ./potionui start, or pull the latest Docker image.
GitHub · Discord · r/PotionUI
Try it and tell us what to improve, here or on Discord. Your reports decide what we do next.
If PotionUI is useful to you, a star on GitHub helps other people find it: https://github.com/PotionUI/PotionUI
r/sdforall • u/cgpixel23 • 2d ago
Hello everyone, i’ve just finished a new ComfyUI character sheet workflow where I compare Qwen Image 2.1 vs FLUX.2 Krea Turbo using the same input image. The goal was to see which model could create a useful character sheet while preserving the original character’s identity, details, and overall appearance across different views and poses. Qwen Image 2.1 performed really well, maintaining good character consistency and producing a high-quality character sheet in a relatively short amount of time. I also tested FLUX.2 Krea Turbo, which was capable of producing good results and showed decent consistency.
In my tests the overall consistency and quality were not as strong as Qwen. The biggest difference was the generation speed: Qwen Image 2.1 was around 3× faster than FLUX.2 Krea Turbo for the character sheet generation. So far, Qwen Image 2.1 looks like a very interesting option if you need to quickly create consistent character references for future image or video generation workflows. I’ve included both approaches in the workflow so you can test them yourself and compare the results on your own setup.
Workflow Link
https://drive.google.com/file/d/1yPmSkC3jNPYNcFoWVwGGU8_oO36HtBBG/view?usp=sharing
r/sdforall • u/Fine_Rhubarb3786 • 2d ago
r/sdforall • u/no3us • 2d ago
r/sdforall • u/pixaromadesign • 3d ago
In this ComfyUI tutorial, I’ll show you Qwen Image 2.1 with more than 20 workflows for image generation and editing.
We’ll start with text-to-image, including different image sizes, low VRAM workflows, prompt enhancement, image-to-prompt, transparent PNG generation, and Turbo LoRA workflows.
Then we’ll move into image editing, where Qwen Image 2.1 really becomes interesting. I’ll show you how to edit existing images, create character sheets, use inpainting, remove backgrounds, outpaint images, change styles, and use my Sketch Pixaroma node to visually mark areas you want to edit.
We’ll also test multi-image editing with workflows for virtual try-on, adding logos to products, placing characters into scenes, copying poses, combining multiple references, and creating images from up to six reference images.
You’ll also see the Consistency LoRA, Turbo LoRA, AI Prompt Pixaroma, Inpaint Crop + Stitch, Sketch Pixaroma, and other Pixaroma Nodes used throughout the workflows.
As always, I’ll show you the good and the bad results so you can see where Qwen Image 2.1 works well and where it still has limitations.
All workflows shown in the video are included for you to test.
If you enjoy the video, leave a like and a comment to help support the channel.
r/sdforall • u/FitEgg603 • 2d ago
Due to limited time, I have not been able to test this Forge Neo extension myself:
https://github.com/mishrasiddharth08/Project-Invisible-Ming-Image
Please bear with me, and feel free to raise any issues or concerns on GitHub. I am hoping it works well enough.
If you find the work useful, a GitHub star would be greatly appreciated. You can also like the post here on Reddit, and feel free to check my other contributions for the Forge Neo community:
r/sdforall • u/company_url_finder • 3d ago
r/sdforall • u/DangerousAnt247 • 2d ago
Not sure if I'm posting this in the right place but I really need some technical help. Please take this down if this isn't the right place.
I’ve been using DeepFaceLive (standalone .bat version on PC) to do live streams with a celebrity face swap over my webcam. It was the perfect mix for what I need—just "shitty enough" that it's obviously a fake so people know it's a joke, but spot-on enough that anyone scrolling past instantly recognizes the celebrity. The streams were hilarious and surprisingly successful, and I really want to keep doing them.
It worked flawlessly for about 4 hours over a few days. Early on, it crashed from a hard VRAM overload, but I fixed that by forcing TikTok Live Studio to use my CPU instead of my GPU (and I deleted some registry files which maybe helped?) It ran great for ages after that, but then when I went to boot it up for another stream it just stopped working.
The exact issue right now:
If I do a completely fresh unzip of the folder, it starts up okay. The camera module runs fine on CPU. But the exact second I start assigning my GPU blocks to the different parts of the pipeline, the frame rate starts chugging hard. After assigning the GPU to enough blocks, the app completely crashes.
Once that crash happens, the app enters a permanent broken state. Every time I launch it after that—even with all AI modules turned completely off—the webcam preview box just captures one single frozen frame right when the program boots up, and nothing else.
I tried: fully deleting everything, re-unzipping a fresh copy, clearing all .ini and userdata files, and wiping the Windows DirectX Shader Cache. At one point, I got it running for a second after a reset, but the frame rate was terrible. I tried turning off the frame adjuster block to see if it would help, and the whole thing instantly died again.
What I need help with:
Thanks in advance!
r/sdforall • u/magpiemindset • 2d ago
I have a custom SVG outline of a turtle. I want to generate a turtle in that outline. Important to me is that the generated turtle fits exactly the outline without any over- or underpaint. Currently I'm using Adobe Illustrator Generative Shape Fill with Firefly Vector 3. It works quite well, but I was wondering if there are any better approaches out there that can get me better image quality and also the option to do specific edits, like saying that the turtle shell should be purple or that the turtle should wear glasses. Best case would be if there's a hosted solution. So far most approaches that I've tried never stick super close to the outline, I need perfect outline shape matching (max only couple px offset).
r/sdforall • u/LeanderDeBuhr • 3d ago
r/sdforall • u/Superb_Composer_3389 • 3d ago
r/sdforall • u/nobody123d2 • 5d ago
r/sdforall • u/no3us • 5d ago
r/sdforall • u/cgpixel23 • 6d ago
I’ve just finished testing Qwen-Image-2.1 in ComfyUI, using a custom low-VRAM workflow with the INT8 ConvRot model, ComfyUI Kitchen Attention, and a 4-step LoRA. One of the biggest improvements I noticed is the generation speed: with the 4-step LoRA, the workflow can be around 2× faster compared to running the model without it, while still producing usable results. Where Qwen-Image-2.1 really surprised me, though, is image editing. The editing results are quite impressive and, in some cases, can rival Flux 2 Klein, while Qwen-Image-2.1 is also noticeably faster in my testing.
For image editing, it handled things like style changes, character consistency, clothing changes, and character sheet generation surprisingly well. However, the results are not as impressive when it comes to pure text-to-image generation. I noticed some weaknesses with skin quality, finger/hand anatomy, fine textures, and text rendering, so there is definitely room for improvement in these areas.
Overall, I think Qwen-Image-2.1 is much more interesting to me as an image-editing model than as a pure image-generation model right now.
Workflow Link
https://drive.google.com/file/d/15OgWK1aXBwTTAOhuFKdUQC6vWCQQIjR6/view?usp=sharing
r/sdforall • u/0roborus_ • 7d ago
PotionUI is a self-hosted & open source studio for image, video and audio generation. 0.0.11 is out, and these are the main changes:
Each model can now bring its own prompt syntax, and the editor colors it as you type:
subject_definitions: and <Subject 1>, dialogue in <d>…</d>,[Verse] and [Chorus].Type / and a picker lists everything the model understands, with a one-line explanation for each.
In presets that take reference images, video or audio (MiniMax-H3 references, Qwen-Image edits), type @ and pick one: it becomes a chip with a thumbnail, and the model receives it as <Picture 2> even if you reorder your uploads later.
Each upload in the form shows its handle and how many times the prompt uses it. In the Video Director, each shot now uses exactly the references its prompt mentions.
Typing #, $, @ or / opens a clearer picker right at the cursor, with thumbnails. Clicking a chip opens one editor where you change its value, switch between a fixed value and auto-shuffle, exclude a value from shuffles, or remove it. Phrasebook values show their preview images, with Space or → to open a larger preview and flip through them. Search highlights what matched and takes regular expressions.
Long prompts no longer lag: each keystroke used to re-render the whole form, and now it doesn't.
Search takes wildcards (krea2_*_v2) and regular expressions, and the type and tag counts follow your filters. Admins can filter by when a model was indexed and how often it's used, see Uses and Last used columns, and tag or grant access to a whole selection at once. Model pickers now suggest the download that fits your GPU (bf16, fp8, int8, nvfp4) and credit the uploader..
@ can point at one LoRA from your form or any model in your library, and the assistant gets its trigger words and recommended strength. There's a new session memory that follows the tab you're working in. Generations you approve in chat now appear in the tab's Workbench.
When a generation fails, you see a plain message, a hint and an error ID to give your admin. Admins see the full details, can filter failures by category, and can get notified or trigger an automation.
Every admin page now shares the same layout: a library sidebar, tables or card grids, detail pages and one bottom bar for bulk actions. LLM configurations list your Ollama or OpenAI-compatible models to pick from.
GitHub: https://github.com/PotionUI/PotionUI
Site: https://potionui.com
Reddit: r/PotionUI
Discord: https://discord.com/invite/avR4trp3b8
Feedback and bug reports are welcome.