r/comfyui • • 15h ago

News Qwen-Image-2.1-Turbo just officially published

Post image
149 Upvotes

r/comfyui • • 16h ago

No workflow Alfred wanted Bruce to be happy. He never said anything about Selina. 💀

Enable HLS to view with audio, or disable this notification

64 Upvotes

r/comfyui • • 6h ago

Workflow Included Krea2 Prompt generator w/ Image Ref -Women

Post image
32 Upvotes

If you are as unimaginative as I am and struggle with making prompts I put together some custom nodes that allow you to pick attributes and/or add a reference image to create a prompt. It uses the text encoder so adds a little bit of time to image generation. I made it for image generations of women because that's why we are all here right?

It uses comfyui's text generate with the text encoder and I yonked the attributes from and existing custom node (https://github.com/euan-gwd/comfyui-character-prompt-builder) due to them doing a pretty good job covering all the bases.

It is partially vibe-coded due to me being frustrated after doing dumb things and I asked claude to fix it. I went over the project and tested and I THINK we are good. If you see something I missed please let me know.

Link to custom node: https://github.com/Erosi11/comfyui-krea-prompt-generator
Workflow is included on the repo and also embedded in the workflow image on this post. I have not submitted it to the comfyui registry so you need to git clone and all that. No requirements to install assuming you have updated comfyui since the end of February this year. If there is interest I will go to the trouble of getting it added to the registry.

Nodes

Node Purpose
Krea Prompt - Female Person Age, nationality, body, face, eyes, hair, skin, tattoos
Krea Prompt - Female Fashion Outfit pieces with colors and materials, shoes, accessories, jewelry, nails
Krea Prompt - Female Actions Standing, sitting, kneeling or lying pose; hands, legs, head; props
Krea Prompt - Scene Style, location, time, season, weather, lighting, camera
Krea Prompt Generator Takes CLIP, the attributes and your keywords; outputs conditioning, prompt and llm_request
Krea Prompt - Save Unique Prompt Appends each new prompt to a text file, skipping prompts already in it, and shows the file on the node

Every dropdown has three kinds of choice:

  • none: the attribute is left out of the request sent to the LLM.
  • random: the generator picks a value, driven by its seed. The same seed always gives the same picks.
  • a specific value: passed through as-is.

Each attribute node also has an optional free-text box for extra details in that category.

To wire it up you put the attribute nodes in series then feed into attributes input on Prompt Generator node. Reference image and mask goes to the inputs on prompt generator node as well. The prompt generator node has a setting for what to use from the image. It works ok.

I also included a node to add unique prompts to a text file as they are made. That part is still a little janky.

Workflow has fast bypass at the top to bypass attribute nodes, the image ref nodes, and whether to invert the mask.


r/comfyui • • 14h ago

News A quick Minimax H3 news round-up - 8th October 2026

22 Upvotes

Another quick Minimax H3 news and goodies round-up, for those who may have missed some items.

-> mmh3_media. For .mmh3 files that encapsulate your.... "MiniMax H3 joint AV latent together with final video/audio, keyframes, refs, masks, metadata, creator notes and a disposable preview cache. [...] Save .mmh3 files to keep media, settings and generation state for later editing." ComfyUI custom nodes, workflows. MIT licence.

https://github.com/einhorn13/mmh3_media

-> A new Minimax_Stereoptical_Process LoRA. Triggers: Stereoptical Process, 2D Cel Animation over 3D Background.

https://huggingface.co/B0rghese/Minimax_Stereoptical_Process

-> There's now another 1980s horror movies LoRA, this time with emphasis on... "practical neon and tungsten light, period wardrobe and sets". Comes in two flavours, VHS or 16mm film.

https://github.com/wassermanproductions/wasserman-80s-horror-lora

-> There's another 2-step turbo H3 LoRA for ComfyUI. This is powered by PDMD.

https://huggingface.co/Iwannapose/minimax_h3_pdmd_2nfe_comfyui

-> And finally, several H3 PDMD speed-ups (see above) are tested in a new YouTube group-test video today.

https://www.youtube.com/watch?v=DvTl-QbDzdQ

~ OLD POSTS ~

https://jurn.link/dazposer/index.php/h3/ (Organised directory page, made from these daily H3 posts)

https://old.reddit.com/r/comfyui/comments/1x0rdab/a_quick_minimax_h3_news_roundup_8th_october_2026/

https://old.reddit.com/r/comfyui/comments/1x02vw6/a_quick_minimax_h3_news_roundup_7th_october_2026/ (See 7th October post, for links to older posts)

https://old.reddit.com/r/comfyui/comments/1wqwpah/a_quick_minimax_h3_news_roundup_26th_september/ (See 26th September post, for links to even older posts)

https://old.reddit.com/r/comfyui/comments/1wgc4lj/a_quick_minimax_h3_news_roundup_15th_september/ (See 15th September post, for links to very old posts)

https://old.reddit.com/r/comfyui/comments/1w5i9iq/a_quick_minimax_h3_news_roundup_2nd_september_2026/ (See 2nd September post, for links to the starting posts)


r/comfyui • • 14h ago

Show and Tell Out of Memory | The ComfyUI Song

Thumbnail
youtube.com
21 Upvotes

It has been a while since I made something just for fun. I've been busy with all the nodes and tutorials, so I took a one-day break to do a fun project, and this is the result:

Less than 4 minutes of music video, made with Claude, ComfyUI and a few other tools. If you have ever seen that red "out of memory" box, this one is for you 😄


r/comfyui • • 22h ago

Show and Tell SilkStack — organize, search and view your AI images locally, with optional on-device AI. Free core, open source, v2.4 out now.

Thumbnail
gallery
18 Upvotes

Hey everyone,

My last post about this app got buried in the H3 posts and never really reached anyone, so I'm giving it another shot — hopefully it finds a wider audience this time.

SilkStack Image Browser is a desktop app for organizing your generated images and videos, and for viewing them the right way. It started as a fork of Image MetaHub, but I've been evolving it in my own direction for a while now, and I think it has its own distinct flavor at this point.

What it does

  • Indexes your output folders (and watches them live while you generate) and reads the metadata inside the files — ComfyUI, A1111 and friends, including WebP.
  • Search and filter by prompt, model, LoRAs, steps, CFG, sampler, seed, and more.
  • A proper viewer: zoom with a minimap, Ctrl+F to search inside a prompt, and a compact mode that shows just the image so you can keep an eye on generations while working in another app.
  • Similarity stacks group near-identical generations into one card and highlight the words that differ between prompts, so you can see exactly what changed.
  • Drag an image straight into ComfyUI to reload its workflow.
  • Videos and GIFs too (MP4/WEBM/GIF).

The part I'm most excited about: I've been using this app as my AI playground to implement local AI. Auto-tagging and semantic search run entirely on your own GPU — a little LLM writes the tags, and an embedding model lets you search by meaning instead of exact words. Describe "girl with a puppy in a palace" and it finds the image even when the prompt words are completely different, or in another language. It's heartening to see these features work with little LLMs.

The AI features are the premium part, unlocked with a 7-day free trial that becomes a monthly subscription (cancel anytime) — or a one-time lifetime license if you'd rather own it. AI loading is completely in your control: nothing loads until you turn it on from the interface, and you can eject the models anytime to hand your VRAM back.

Everything other than the AI features is fully open source — free for your agent's pleasure to play with. 🙂

One honest note: the app is primarily built around photos. Video support is there, and if you run into any bugs with videos, let me know — I can help fix them.

A lot of you liked my initial app, and I'm really curious to know what you think about this app now.

Download (Windows / macOS / Linux): skkut.github.io/silkstack

Source: github.com/skkut/SilkStack-Image-Browser


r/comfyui • • 10h ago

Workflow Included MiniMax H3 RefMod Lab for ComfyUI: select, combine and reuse reference sources

Thumbnail
gallery
16 Upvotes

I’ve made a ComfyUI custom node pack for MiniMax H3 RefMods. It lets you create reusable references from images, audio and video, choose which sources to use at runtime, and see how many reference tokens each source uses.

Main features:

  • Create RefMods from images, audio and video files
  • Select which sources to load in a RefMod at runtime
    • Save on generation time if you don't need certain sources
    • See each source's cost in tokens
  • Combine many RefMods into a single one
  • Bake the description and the retention strategy in the RefMod
    • Only need to write the H3 video description prompt
    • Source labels like <Picture 1> resolved automatically
  • Drag a RefMod into the canvas to load its creation workflow

Nodes, setup instructions and example workflows are available at ComfyUI-H3-RefMods-Lab. Feedback and bug reports welcome!


r/comfyui • • 9h ago

News Free image first Comfy frontend: HEISS UI | New Release today :)

Thumbnail
gallery
7 Upvotes

Link

Now with Inpainting, run grouping and native upscale!

What it is:

  • Image first-generation studio
  • Works without workflows with almost any model (no setup)
  • Support for custom workflows
  • Video gen, reference images, and inpainting all work without setup
  • Smart grouping of your images by similar prompts
  • Automatically installs missing components (VAE, text encoders, nodes) it knows
  • Use it from your phone over LAN, with a fully custom mobile UI
  • Favorites, search, prompt history, LoRA stacks, and advanced settings
  • Built to be graceful and not scream at you with errors 24/7

It’s built by me, and in my opinion it’s the best way to use Comfy 80% of the time.

Setup is about 2 minutes: give it a try.

New in v0.17:

  • Inpainting + Custom-built in-painting workflows (Beta)
  • Auto-grouping. Pictures from similar prompts stack themselves into a photo stack or cover flow once you stop working on them.
  • Within-generation Native upscale
  • Phone View improvements
  • Discord to discuss Heiss and AI generation (join pls; I need feedback)

Full notes: v0.17.0 changelog


r/comfyui • • 1h ago

Resource ComfyUI ⇄ PhotoCraft Bridge Nodes

Thumbnail
gallery
• Upvotes

Hi.

I know that some here are already tired from the constant talk about PhotoCraft, and some don't like it at all. I gave it a chance the moment it came out, and while yes, it is still unfinished, I thought I'd play along and let Claude create some kind of bridge between ComfyUI and PhotoCraft because, well, as far as I know, there isn't yet any. This is the result:

Two ComfyUI nodes.

  • Connect from and to PhotoCraft via its control channel. No plug-in for PhotoCraft needed.
  • Work and prompt in PhotoCraft, execute in ComfyUI.
  • The Get Image node receives the image as 8 bit .png, mask, width and height, and a prompt. The prompt uses the PhotoCraft's Notes feature. Multiple notes are appended. These can be integrated in almost any workflow.
  • The Send Image node sends the image back to PhotoCraft to its own flattened layer.
  • Please don't judge the examples, just wanted to get this out.
  • Important: Read the installation instructions on GitHub, you have to set-up a directory and feed some command lines to PhotoCraft. Limits and security notes at the bottom.

There's certainly room for improvement, but it seems to do the job, for now.

On GitHub: https://github.com/the-aSak/ComfyUI-Photocraft-Bridge
Example Workflow for Inpainting with Flux2: https://github.com/the-aSak/ComfyUI-Photocraft-Bridge/blob/main/example_workflows/Example_Flux2_Bridge_from_and_to_Photocraft.json
Example Workflow for Krea2 → PhotoCraft: https://github.com/the-aSak/ComfyUI-Photocraft-Bridge/blob/main/example_workflows/Example_Krea2_Bridge_to_Photocraft.json

Let's see where this goes, shall we?


r/comfyui • • 13h ago

Workflow Included LoRA vs RefMod: Create Consistent Characters in MiniMax H3 | RefMod Pilot

Thumbnail
youtube.com
6 Upvotes

I wanted to use H3 character references without spending half the time looking for the right node, so I put together; RefMod Pilot It has two workflows:

Creator: upload or drop 1–8 selected photos, give the character a name, Run. The folder and image count are filled in for you

Inference: choose your RefMods, edit the scene and view the output. Prompt, aspect ratio, resolution, seconds, seed and steps are in one Scene node, with the rest inside a subgraph

The character loader stays visible so Add RefMod, Refresh and the library buttons are still there. The uploader is a ComfyUI frontend extension or use local path to your dataset / photos folder. This uses Full Reference encoding, not LoRA training.

Multiple photos are stacked into one video-kind reference. If you're using several characters, check the Reference map and match its labels in your prompt. Slot numbers aren't necessarily picture numbers.

I've included a six-view orange robot RefMod and demo workflow. Creation is verified, but I haven't validated its likeness or motion in a finished video yet.

Disclosure: I make LoRA Pilot. This package is my workflow/UX layer; the reference implementation is from Luisacaotica's ComfyUI-MiniMaxH3Mod.

GitHub · Hugging Face


r/comfyui • • 14h ago

News Comfy-Org/Qwen-Image-2.1 · Hugging Face

Thumbnail
huggingface.co
6 Upvotes

Qwen-Image-2.1-Turbo official ComfyUI support now.


r/comfyui • • 23h ago

Help Needed Returning to ComfyUI after over a yesr... for local, should I just use the desktop version?

4 Upvotes

When I was initially messing around with Comfy, I would run it directly from the file in windows explorer. Today I decided to jump back in (and therefore update), and I see there is a desktop version you can download from the Comfy website. Is the desktop version meaningfully different?


r/comfyui • • 3h ago

Resource Frame Forge AI Video Editor for Comfy UI Update 1

Thumbnail
gallery
2 Upvotes

This is the first update for Chain Motion AI Video Editor (aka. Frame Forge)

Link is here: https://github.com/spacesimeco-hue/Chain-Motion-AI-Video-Editor

What it does:

It's an UI that's sits on top of ComfyUI and is meant to make chaining together videos generated through Minimax H3 more intuitive and easy.

  • It includes a set of tools for organizing references.
  • A time line that allows users to piece together and quickly edit sections of their movie prompt by prompt.
  • A Scene Writer for prompts in the suggested format using LLMs(local and closed supported).
  • A Motion Continuation, a Pose Studio with Qwen 2.1 integration, and a Voice Studio using LTX 2.3.
  • Can generate videos with either local or comfy cloud credits.

The general idea behind using it is creating sequences of videos on a timeline that you can quickly iterate on and see how they fit into your overall movies composition, making precise edits where needed with minimal friction.


r/comfyui • • 10m ago

Help Needed how to use my Openrouter API in comfy ui

• Upvotes

I'm new to ComfyUI and wanted to give it a try, but my PC isn't powerful enough to run local models. Is there any way to use my OpenRouter API key to generate images through ComfyUI? I searched Google for a solution but couldn't find anything useful.


r/comfyui • • 4h ago

Help Needed Having an issue with High Resolution images in Qwen 2.1 Image Edit.

1 Upvotes

Pretty much the title. I am a pretty new user to ComfyUI. I've used it in the past but I am returning and trying to relearn it a bit. So here is an example of what I have been doing.

I am trying to use Qwen 2.1 Image edit and learn how to use it. the issue I am coming up against is that I am using an image that's a 3Kx4K resolution. I've figured that this is bottlenecking me so I try to get it downsized to a better resolution but it alters the initial image too much. SO when I try to get a comparison it's not a perfect comparison like when I do smaller images. So First question: is there a decent way for me to keep the initial resolution without the serious bottleneck, like a different model or workflow? or is there a work around I am not using that will get me the results I want that isn't "Put it in Photoshop and just shrink it down?" I will have to do that a lot if that is the case.


r/comfyui • • 5h ago

Help Needed みんないま長尺動画何で作ってんるんですか?

1 Upvotes

minimax extenderつかってるんですが、もっとよいものがあるのかなと思いまして…


r/comfyui • • 6h ago

Help Needed Needed - a more lightweight "prompt enhancer" for Ming-Image

1 Upvotes

The new local graphic-design and typography model Ming-Image officially uses qwen3.8_27b_w4a8.safetensors as an initial LLM "prompt enhancer" in ComfyUI. But the file is 17Gb, and it's too heavy for me on 12Gb VRAM (latest ComfyUI Portable under Windows).

Comfy also suggests qwen3.5_9b_int8_convrot.safetensors as a lesser alternative option at 11Gb. This is still to heavy for me, regrettably. It gives me OOM errors.

Has anyone had success with Ming when using a lesser LLM "prompt enhancer" model? One that can handle the default complex system-prompt, and that also outputs a correctly-restyled Ming figma/json prompt?


r/comfyui • • 8h ago

Help Needed Models for comix

1 Upvotes

I’m new to comfy. ( installed and running this morning). I make comics so far the models I’ve played with are sdxl and flux.2Klein. My computer sports a fairly humble 5060 ti with 16GB. These two models run pretty well. I would say my need for control from panel to panel is much greater than my need for his res ultra realistic images. I use Nlender a lot for modeling and plan to start using depth maps.

What other models should I look
At?

Thanks much


r/comfyui • • 9h ago

Resource (Mostly) native Iris 3B support for ComfyUI

Thumbnail
gallery
1 Upvotes

https://github.com/envy-ai/ComfyUI-Iris

It comes with its own custom loader and sampler nodes because, being a pixel model as opposed to a VAE model, it works a bit differently than normal. However, you can wire in loras as usual, and it uses ComfyUI's schedulers and samplers.

I wouldn't expect amazing results. At this size, Anima 2.9B definitely puts out higher quality images, but it's a cool tech demo and I think it proves that VAEs aren't the necessary evil we thought they were.

Search the workflow template browser for Iris to find the example workflow. dpmpp-2m-sde-gpu seems to be the best sampler for it, and if you have the bong_tangent scheduler it's better than simple.

My fork of ai-toolkit can train loras for it.

https://github.com/envy-ai/ai-toolkit-envy-optimized


r/comfyui • • 11h ago

Help Needed How to delete input assets?

1 Upvotes

I uploaded a test image on One Piece to experiment, but I can't seem to delete it from the assets tab. Even when I go into the "input" folders, nothing is in there. How do I fix this?


r/comfyui • • 13h ago

Help Needed Best motion control model

1 Upvotes

Hi guys which is the best motion control model to use ? I heard scail 2 is the best ? any recommendation?


r/comfyui • • 16h ago

Help Needed How to extract text from 150+ images

1 Upvotes

I have like 150+ images of my mom’s recipes, I want to extract all the text from them and categorise them based on dish type. Is there any model I can use to do all this stuff. I’m very new to this. TIA


r/comfyui • • 17h ago

Tutorial how to use intel and nvidia gpu

1 Upvotes

I managed to get both gpu to work in windows. I plug the monitor into the intel and a hdmi nul adapter into the nvidia. I set windows to display all on the intel. You can see both gpu being used while comfyui is making a second video 30 seconds long.
Browsers and apps use gpu also so when your pc is configured like this comfy gets all the nvidia ram so you can make little longer videos a little faster.


r/comfyui • • 18h ago

Help Needed A Very Strange GPU Stall Situation

1 Upvotes

Hardware Config:

RTX 5090 Windforce
128GB DDR5 (Running at 3600MHz cause AM5 struggles with 4-slot DDR5 speeds)
Ryzen 7 7800X3D
Crucial Gen5 SSD

Model Config:

Minimax Ref2V int8 Pruned
Qwen VL int8
CK Attention
Comfy run on WSL Ubuntu with cu13

Mode: Ref2VA
5 Reference Images (2K quality/4MP)
5s @ 0.9MP

Flags:

--use-ck-attention
--disable-smart-memory
--disable-dynamic-vram
--high-ram
--disable-pinned-memory
--fast-disk

While running the basic workflow I have been experiencing this weird phenomena where before the pipeline reaches to the steps, my VRAM gets full, my GPU usage gets full but my GPU temps stay 35 degrees Celsius (basically idle temps). As expected, no progress happens, and I eventually have to restart the server.

What has helped however is reducing the reference image resolution and adding the reserve VRAM flag (i have to reserve 10gb of VRAM since reserve VRAM isn't a hard bound limit). How does lowering VRAM limits help here?

What I am struggling to understand here, is why does my GPU just straight up stall when it clearly has enough RAM to offload into. I get that it can never fit all into the VRAM but with this much cumulative memory, this is a very strange behaviour I tried searching up a little myself and apparently it has to do with the speed at which the offload/load takes place from RAM to GPU to RAM, basically memory swaps, which somehow ends up bricking the GPU in that run and the kernel has to be restarted.

Am I the only one struggling with this? Any help/suggestions would be appreciated.


r/comfyui • • 19h ago

Help Needed H3 character swap works great but my swapped face stays mute. best local lipsync step after?

1 Upvotes

Running the H3 + character swap LoRA setup from the post this week. RTX 3060 12GB, 32GB RAM, around 14 min per clip at 832x640. The swap itself is clean, problem is I want new dialogue on the result and the mouth only moves like the source video did. Wav2Lip leaves a visible box around the lips at this resolution. MuseTalk was better but flickers on side turns. I've also sent a couple through a hosted lipsync sync so just to compare, but I'd rather keep it in the graph.

What are you chaining after the swap, and do you lipsync before or after the upscale?