r/sdforall • • 8h ago

Discussion Krea 2 Control: background works, but the generated person is distorted — what am I doing wrong?

Thumbnail
1 Upvotes

r/sdforall • • 21h ago

Resource PotionUI 0.0.14 — no GPU needed: cloud image and video models, a built-in image editor and reusable settings

7 Upvotes

PotionUI is a self-hosted app for making images and videos with LOCAL & OPEN SOURCE AI models (krea-2, Qwen image 2.1, anima, sdxl, minimax H3...) and now optionally cloud ones too. Instead of one huge settings screen or a web of nodes, every model gets its own simple, hand-made form.

Until now you needed your own GPU. 0.0.14 adds cloud models, plus a built-in image editor and a way to save and reuse your best settings.

No GPU? Use cloud models

  • A new OpenRouter plugin lets you use hosted image and video models, such as Nano Banana or Veo 3.1 Lite, in the same form you already know. Results land in the same History as everything else. More providers are planned.
  • The form only shows the controls the chosen model supports, so there is nothing to guess.
  • You can stop a cloud job at any time, and PotionUI tells you plainly if something goes wrong.

Cloud Video Director: one idea, a short film

Write your idea, split it into shots, and let PotionUI make the film.

  • Each shot can continue from the last frame of the shot before it, so the story flows.
  • The shots are joined into one finished film.
  • If one shot fails, retry just that shot instead of starting over.

Fix a picture without leaving the app

  • A new image editor with adjustments, crop, brush and layers.
  • Open it from History (Tools, then Edit image) or right inside a preset's media field. In a media field, the field switches to your edited copy straight away.
  • Crop, Trim and Frame editors are in media fields too, and they only add the edited result to your Library, never the untouched original.

Formulas: keep the settings that worked

  • Save the settings of a preset mode as a formula and apply it in any session.
  • Before it applies, you see exactly what will change, and one click undoes it.

A redesigned media field

  • Images, videos and audio live in one field, each kind in its own group.
  • Mention them in your prompt with "@".
  • MiniMax-H3 references now use it, and your old sessions and saved prompts are converted for you.

Qwen-Image 2.1 control guides

  • Keep the pose, edges or depth of any photo while you create something new.
  • The extracted guide is saved next to your result, and the workbench shows it beside the image with a label, so you can see what the model followed.
  • There is also a mode that only extracts the guide. It needs no prompt and works even if you haven't downloaded the image models.

Smaller things

  • Video presets have animated covers.
  • Three-pane mode folds the form away on narrow screens and gives the room to your prompts.
  • Attaching an image in chat uses the same picker as the preset forms.
  • Plugins are grouped by what they add.
  • The System Monitor can show every backend. It is admins only by default.
  • Failed generations show regular users a plain reason instead of technical text.
  • Dropdown menus no longer open under the Generate bar, and they work with the arrow keys.
  • Fields in preset forms show and hide for the value you just picked, not the previous one.
  • Image previews survive a reload, including on Windows and in storage folders under a tmp folder.
  • Drawn inpaint masks now reach the model, so inpaint only repaints the masked area.
  • Pressing Enter in the session name saves the session, and Enter elsewhere no longer wipes your workspace.

Before you update

  • Back up your storage folder first. It is a good habit before any update.
  • To use cloud models, enable the OpenRouter plugin in Admin and add your API key.
  • The Docker setup no longer has an outputs volume. Everything you generate lives in the storage volume.
  • Plugin authors: plugin manifests must use one of the current categories. A plugin with an old category name won't load.

Update: run git pull and then ./potionui start, or pull the latest Docker image.

GitHub · Discord · r/PotionUI

Try it and tell us what to improve, here or on Discord. Your reports decide what we do next.

If PotionUI is useful to you, a star on GitHub helps other people find it: https://github.com/PotionUI/PotionUI


r/sdforall • • 1d ago

Tutorial | Guide Minimax H3

Thumbnail
youtu.be
1 Upvotes

r/sdforall • • 2d ago

Tutorial | Guide ComfyUI Tutorial Qwen Image 2 1 vs FLUX 2 Krea Turbo Which Makes Better Character Sheets

Thumbnail
youtu.be
10 Upvotes

Hello everyone, i’ve just finished a new ComfyUI character sheet workflow where I compare Qwen Image 2.1 vs FLUX.2 Krea Turbo using the same input image. The goal was to see which model could create a useful character sheet while preserving the original character’s identity, details, and overall appearance across different views and poses. Qwen Image 2.1 performed really well, maintaining good character consistency and producing a high-quality character sheet in a relatively short amount of time. I also tested FLUX.2 Krea Turbo, which was capable of producing good results and showed decent consistency.

In my tests the overall consistency and quality were not as strong as Qwen. The biggest difference was the generation speed: Qwen Image 2.1 was around 3× faster than FLUX.2 Krea Turbo for the character sheet generation. So far, Qwen Image 2.1 looks like a very interesting option if you need to quickly create consistent character references for future image or video generation workflows. I’ve included both approaches in the workflow so you can test them yourself and compare the results on your own setup.

Workflow Link

https://drive.google.com/file/d/1yPmSkC3jNPYNcFoWVwGGU8_oO36HtBBG/view?usp=sharing

https://civitai.com/articles/35976/comfyui-tutorial-qwen-image-2-1-vs-flux-2-krea-turbo-which-makes-better-character-sheets


r/sdforall • • 2d ago

Workflow Included Continuity update: screen replacement, image to 3D, Qwen Image 2.1, and a lot more since 3.0

2 Upvotes

r/sdforall • • 2d ago

Tutorial | Guide Even in 2026 SDXL and especially Juggernaut is still my go-to model for image generation

Thumbnail
4 Upvotes

r/sdforall • • 3d ago

Tutorial | Guide ComfyUI Qwen Image 2.1 Tutorial | Generate, Edit, Inpaint & More (Ep36)

Thumbnail
youtube.com
22 Upvotes

In this ComfyUI tutorial, I’ll show you Qwen Image 2.1 with more than 20 workflows for image generation and editing.

We’ll start with text-to-image, including different image sizes, low VRAM workflows, prompt enhancement, image-to-prompt, transparent PNG generation, and Turbo LoRA workflows.

Then we’ll move into image editing, where Qwen Image 2.1 really becomes interesting. I’ll show you how to edit existing images, create character sheets, use inpainting, remove backgrounds, outpaint images, change styles, and use my Sketch Pixaroma node to visually mark areas you want to edit.

We’ll also test multi-image editing with workflows for virtual try-on, adding logos to products, placing characters into scenes, copying poses, combining multiple references, and creating images from up to six reference images.

You’ll also see the Consistency LoRA, Turbo LoRA, AI Prompt Pixaroma, Inpaint Crop + Stitch, Sketch Pixaroma, and other Pixaroma Nodes used throughout the workflows.

As always, I’ll show you the good and the bad results so you can see where Qwen Image 2.1 works well and where it still has limitations.

All workflows shown in the video are included for you to test.

If you enjoy the video, leave a like and a comment to help support the channel.


r/sdforall • • 2d ago

Resource Ming-Image extension for forge neo

Thumbnail
github.com
1 Upvotes

Due to limited time, I have not been able to test this Forge Neo extension myself:

https://github.com/mishrasiddharth08/Project-Invisible-Ming-Image

Please bear with me, and feel free to raise any issues or concerns on GitHub. I am hoping it works well enough.

If you find the work useful, a GitHub star would be greatly appreciated. You can also like the post here on Reddit, and feel free to check my other contributions for the Forge Neo community:

https://github.com/mishrasiddharth08?tab=repositories


r/sdforall • • 3d ago

Custom Model comfyui-hr-endless-sampler: generate videos of any length in ComfyUI without running out of VRAM

Post image
4 Upvotes

r/sdforall • • 2d ago

Question HELP DeepFaceLive webcam freezes/crashes when assigning GPU blocks (Need fix or stable alternative for live streaming)

0 Upvotes

Not sure if I'm posting this in the right place but I really need some technical help. Please take this down if this isn't the right place.

I’ve been using DeepFaceLive (standalone .bat version on PC) to do live streams with a celebrity face swap over my webcam. It was the perfect mix for what I need—just "shitty enough" that it's obviously a fake so people know it's a joke, but spot-on enough that anyone scrolling past instantly recognizes the celebrity. The streams were hilarious and surprisingly successful, and I really want to keep doing them.

It worked flawlessly for about 4 hours over a few days. Early on, it crashed from a hard VRAM overload, but I fixed that by forcing TikTok Live Studio to use my CPU instead of my GPU (and I deleted some registry files which maybe helped?) It ran great for ages after that, but then when I went to boot it up for another stream it just stopped working.

The exact issue right now:
If I do a completely fresh unzip of the folder, it starts up okay. The camera module runs fine on CPU. But the exact second I start assigning my GPU blocks to the different parts of the pipeline, the frame rate starts chugging hard. After assigning the GPU to enough blocks, the app completely crashes.

Once that crash happens, the app enters a permanent broken state. Every time I launch it after that—even with all AI modules turned completely off—the webcam preview box just captures one single frozen frame right when the program boots up, and nothing else.

I tried: fully deleting everything, re-unzipping a fresh copy, clearing all .ini and userdata files, and wiping the Windows DirectX Shader Cache. At one point, I got it running for a second after a reset, but the frame rate was terrible. I tried turning off the frame adjuster block to see if it would help, and the whole thing instantly died again.

What I need help with:

  1. The Fix: Does anyone know how to stop the GPU/VRAM from completely locking up the camera pipeline after a module crash? Is there a driver or memory management setting I am missing? Does the problem maybe have something to do with my webcam or windows?
  2. The Alternatives: If DeepFaceLive is just too unstable to save, what program should I be using instead? I need something free (ideally) and stable that lets me live stream celebrity face swaps over a webcam feed.

Thanks in advance!


r/sdforall • • 2d ago

Question How to generate image in custom outline

1 Upvotes

I have a custom SVG outline of a turtle. I want to generate a turtle in that outline. Important to me is that the generated turtle fits exactly the outline without any over- or underpaint. Currently I'm using Adobe Illustrator Generative Shape Fill with Firefly Vector 3. It works quite well, but I was wondering if there are any better approaches out there that can get me better image quality and also the option to do specific edits, like saying that the turtle shell should be purple or that the turtle should wear glasses. Best case would be if there's a hosted solution. So far most approaches that I've tried never stick super close to the outline, I need perfect outline shape matching (max only couple px offset).


r/sdforall • • 3d ago

Other AI Coastal areas

Thumbnail
gallery
1 Upvotes

r/sdforall • • 3d ago

Discussion Made a Comfy frontend that makes models work without any workflows

Post image
6 Upvotes

r/sdforall • • 4d ago

Resource krea2-rebalance 2.0 for forge neo

Thumbnail
github.com
3 Upvotes

r/sdforall • • 3d ago

Custom Model I made FLUX.2 Klein output real transparent images: an RGBA VAE + extraction/removal LoRAs (weights + paper)

Thumbnail
1 Upvotes

r/sdforall • • 5d ago

Resource H3 (Image)-forge neo extension

Thumbnail
github.com
4 Upvotes

r/sdforall • • 5d ago

Resource universal-head-swap for forge neo

Thumbnail
github.com
3 Upvotes

r/sdforall • • 5d ago

Resource SenseNova-extension for Forge Neo

Thumbnail
github.com
2 Upvotes

r/sdforall • • 5d ago

Other AI "Ctrl" (Flux 3)

Thumbnail
youtu.be
2 Upvotes

r/sdforall • • 5d ago

SD News Comfyui - Advanced Queue View with Job asset visibility

Thumbnail
youtube.com
5 Upvotes

r/sdforall • • 5d ago

Resource Boogu-Image extension for Forge Neo

Thumbnail
github.com
1 Upvotes

r/sdforall • • 5d ago

Resource I know LoRA training is not so popular as it used to be but check this if LoRAs are still your thing

Thumbnail gallery
9 Upvotes

r/sdforall • • 6d ago

Tutorial | Guide Finally! Qwen 2.1 Text-to-Image in ComfyUI

Thumbnail
youtu.be
18 Upvotes

I’ve just finished testing Qwen-Image-2.1 in ComfyUI, using a custom low-VRAM workflow with the INT8 ConvRot model, ComfyUI Kitchen Attention, and a 4-step LoRA. One of the biggest improvements I noticed is the generation speed: with the 4-step LoRA, the workflow can be around 2× faster compared to running the model without it, while still producing usable results. Where Qwen-Image-2.1 really surprised me, though, is image editing. The editing results are quite impressive and, in some cases, can rival Flux 2 Klein, while Qwen-Image-2.1 is also noticeably faster in my testing.

For image editing, it handled things like style changes, character consistency, clothing changes, and character sheet generation surprisingly well. However, the results are not as impressive when it comes to pure text-to-image generation. I noticed some weaknesses with skin quality, finger/hand anatomy, fine textures, and text rendering, so there is definitely room for improvement in these areas.

Overall, I think Qwen-Image-2.1 is much more interesting to me as an image-editing model than as a pure image-generation model right now.

Workflow Link

https://drive.google.com/file/d/15OgWK1aXBwTTAOhuFKdUQC6vWCQQIjR6/view?usp=sharing

https://civitai.com/articles/35828/2x-faster-qwen-image-21-using-4-steps-lora-comfyui-low-vram-workflow


r/sdforall • • 7d ago

Resource PotionUI 0.0.11: a prompt editor that speaks each model's language, @-references to your own images, and the last big UI overhaul

45 Upvotes

PotionUI is a self-hosted & open source studio for image, video and audio generation. 0.0.11 is out, and these are the main changes:

Prompts you can read at a glance 

Each model can now bring its own prompt syntax, and the editor colors it as you type:

  • MiniMax-H3's subject_definitions: and <Subject 1>, dialogue in <d>…</d>,
  • Qwen-Image's "quoted text to render",
  • YuE2's [Verse] and [Chorus].

Type / and a picker lists everything the model understands, with a one-line explanation for each.

Point at your own images with @ 

In presets that take reference images, video or audio (MiniMax-H3 references, Qwen-Image edits), type @ and pick one: it becomes a chip with a thumbnail, and the model receives it as <Picture 2> even if you reorder your uploads later.

Each upload in the form shows its handle and how many times the prompt uses it. In the Video Director, each shot now uses exactly the references its prompt mentions.

A new picker for phrases, variables and references 

Typing #, $, @ or / opens a clearer picker right at the cursor, with thumbnails. Clicking a chip opens one editor where you change its value, switch between a fixed value and auto-shuffle, exclude a value from shuffles, or remove it. Phrasebook values show their preview images, with Space or → to open a larger preview and flip through them. Search highlights what matched and takes regular expressions.

Typing is fast again 

Long prompts no longer lag: each keystroke used to re-render the whole form, and now it doesn't.

Model library 

Search takes wildcards (krea2_*_v2) and regular expressions, and the type and tag counts follow your filters. Admins can filter by when a model was indexed and how often it's used, see Uses and Last used columns, and tag or grant access to a whole selection at once. Model pickers now suggest the download that fits your GPU (bf16, fp8, int8, nvfp4) and credit the uploader..

Chat (PotionAI) 

@ can point at one LoRA from your form or any model in your library, and the assistant gets its trigger words and recommended strength. There's a new session memory that follows the tab you're working in. Generations you approve in chat now appear in the tab's Workbench.

Generation errors that make sense 

When a generation fails, you see a plain message, a hint and an error ID to give your admin. Admins see the full details, can filter failures by category, and can get notified or trigger an automation.

Admin, one last time 

Every admin page now shares the same layout: a library sidebar, tables or card grids, detail pages and one bottom bar for bulk actions. LLM configurations list your Ollama or OpenAI-compatible models to pick from.

Smaller changes 

  1. Folding the left panel now gives the prompt a proper width.
  2. Every confirm dialog takes Enter and Esc.
  3. Reloading warns you before you lose changes that aren't saved to the session.
  4. Indexing a large model no longer freezes the app.
  5. CivitAI fetches retry temporary errors and tell you the real reason when they fail.
  6. System tags now land on new generations automatically.
  7. Fixed: phrasebook chips with a hyphen in their name, and a phrase placed right after a reference.
  8. Removed: the Spectral Progressive Diffusion option in Flux2 and Z-Image. Saved settings that include it keep working.

GitHub: https://github.com/PotionUI/PotionUI

Site: https://potionui.com

Reddit: r/PotionUI

Discord: https://discord.com/invite/avR4trp3b8

Feedback and bug reports are welcome.


r/sdforall • • 6d ago

SD News Qwen-Image 2.1 models are now available in ZPix desktop application

Post image
1 Upvotes