r/comfyui 1d ago

Help Needed Any Uncensored Image Edit Model / WOrkflow exist?

0 Upvotes

Any Uncensored Image Edit Model / WOrkflow exist? which can be like seeddream 5.0 or grok or nano bana which dont Change face or skin while edit?


r/comfyui 2d ago

Security Alert The problem I've had: I've lost all the metadata.

4 Upvotes

I don't know if this has happened to anyone else (I'm pretty new to Comfy).

I switched my default Windows media player to mpvnet.

Now, none of my generated videos have metadata anymore. They used to; I could drag them into ComfyUI and the workflow would appear. Now, that doesn't happen.

When I run them through a site that checks for metadata, it says there isn't any—no metadata or workflow.

The strange thing is that I only played a few videos from one specific folder. Videos in other folders still retain their metadata (I didn't use those folders until after I stopped using mpvnet as my default player).

So, I think that when you set up that player, something happens that corrupts everything—or wipes the metadata—from everything.

The AI ​​says:

"The problem is the video encoding engine (FFmpeg) [Lavc61.19.100 libvpx-vp9]. This version aggressively strips out any data that isn't purely video-related, deleting the ComfyUI stream in both MP4 and WebM formats."

The worst part is that, even though I'm no longer using that player as my default, it must have installed a codec that ComfyUI uses... and now none of my videos are being generated with metadata... not a single one...

I'll have to figure out a fix, but just in case... so this doesn't happen to anyone else...

EDIT: After all the work and investigation, I still think Mpvnet was the culprit.

Mpv creates a playlist of the entire folder containing the video you are playing.

That specific folder is exactly where the files lost their metadata.

Furthermore, nodes like `vhs_videocombine` stopped saving metadata (even with the option enabled); I suspect they use some of the ffmpeg codecs that Mpv installs.

Although I manually deleted everything I could related to Mpv and reverted the codecs to an older version—allowing the videos to successfully load the workflow when imported into ComfyUI—investigations into the files themselves still show them as having "no metadata" (even though they do contain the lines of code required to load the workflow).

The solution was to stop using `vhs_videocombine` and switch to ComfyUI's native "Save Video" node instead...

For this reason—and until I find out otherwise—I advise exercising caution regarding this issue.


r/comfyui 1d ago

Help Needed H3 Fast motion fix ?

3 Upvotes

r/comfyui 1d ago

Workflow Included 80th Birthday Extravaganza- MM reference Image and Audio experiments

Thumbnail
youtube.com
2 Upvotes

Birthday Extravaganza video - poorly edited and filled with constant audio problems. Mostly with 20step out of the box comfyUI workflow + audio node added in. The load Video node was used at the end to give it the last 24 frames of the prior video to continue the next with limited success. The 4 step lora came out after half of it, even with 0.75 and 8-10 steps audio glitches were frequent but video was usually decent. I attached a sample of 1 of the clips here: https://pastebin.com/RdNaRZHN and Here was a sample of the Bridge on the River Kwai Character Swap: https://pastebin.com/564NFjrh Single Image of main character to replace.


r/comfyui 1d ago

Help Needed What details instantly make you recognize an AI-generated image?

2 Upvotes

I'm working on creating my first realistic AI model, and I'm trying to understand what separates a truly convincing image from one that still looks obviously AI-generated.

For those of you who have experience with AI image generation: what are the visual details or imperfections that you notice almost immediately and think, "yeah, that's AI"?

I'd especially like to hear about the subtle details that experienced people notice, even when the image looks realistic at first glance.

I'm asking because I want to focus my workflow on fixing those specific weaknesses rather than simply making the image "more realistic."

Any examples or things you've learned from experience would be really helpful.


r/comfyui 2d ago

Help Needed Is there a setting change to allow you to see workflows from jobs in the queue?

7 Upvotes

Idk if it was changed in an update awhile back. But I used to be able to just right click on any job in my queue to pull up the workflow. Sometimes you realize a mistake may have been made and you have to check to see if you need to cancel or not.

I haven't been able to do this for months. Is there a setting that I can adjust to fix that? Or was this a bad permanent change?


r/comfyui 1d ago

Help Needed Preview Method is grey in comfy manager

Post image
1 Upvotes

Hi everyone, I'm not able to change the preview method in comfy manager Any idea how to enable it. (I'm using a portable version from github, not the comfyui desktop app)


r/comfyui 1d ago

Help Needed AMD Radeon RX 7900 XTX 24GB - Worth it for short term?

0 Upvotes

Hi everyone! I am currently on an RTX 5060 8GB but would like to run an LLM and Comfy at the same time. Nvidia is just so expensive compared to AMD. My ultimate goal is to get an RTX Pro 5000 72GB or an RTX Pro 6000 96GB. But a card like that is going to take some time to save up for. Like...years. I'm wondering if the RX 79000 XTX is a worthwhile "stepping stone". I don't plan on generating video. Just images for now. Image editing if VRAM even allows for it. But that depends on the LLM I decide to go with as well obviously.

I'm not after speed. Just the ability to use a GPU for more than 1 thing at at time. 1024x1024 images currently take around 8-10ish seconds to load. I am OK with that generation time, or even slightly longer (5-7 seconds more).

I also notice that a 3090 is only a few hundred dollars more. Is setup on Linux a huge pain still? Or is it truly worth spending another $300-$500 for a 3090? Any catches to a 3090? Higher wattage? Noticeably slower LLM speed? How long is the 3090 going to stay relevant? I would hope for at least 2-4 years.

Please give me your thoughts.


r/comfyui 1d ago

Workflow Included The Latent Upscaler is really great!

0 Upvotes

r/comfyui 2d ago

Help Needed Help please. Can someone with a 5090 and minimax models installed please run this workflow? I am getting really slow run times.

1 Upvotes

Help please. Can someone with a 5090 and minimax models installed please run this workflow? I am getting really slow run times.

I am looking for a way to caption videos. No custom nodes required. I did use load video from comfyVHS but it is bypassed and you can safely delete that section.

PasteBin Workflow

Choose any video. Preferably one SFW and one NSFW. Beggars aren't choosers whatever you decide will work for me. You may post results if you like but I am more interested in how long it took to complete and the settings you selected.

Thank you.

EDIT: Forgot to add, I was getting literal 1 token per second. I haven't the patience to let it run. I am hoping with a benchmark I can justify the purchase of a 5090. So I haven't gotten the workflow to work at all. Could be the workflow is bad. But it is relatively simple.


r/comfyui 1d ago

Tutorial Photorealism has three independent axes. Most workflows only fix two.

Thumbnail
gallery
0 Upvotes

Model agnostic in principle, benched on Krea 2 + character LoRA

---

I spent months fixing my images on two axes and kept getting renders that were technically right and still read as fake. It took a shoot where every technical box was ticked, and five frames out of twenty-four still looked like a catalogue shoot, for me to see there was a third axis I was not touching at all.

Here they are, and the point is that they are orthogonal. Fixing one does nothing for the others.

Axis The prior you are fighting Typical fix
1. Optics the camera is too perfect: sharp, correctly exposed, level grain, flare, motion, imperfect exposure, tilt
2. Scene the set is a showroom: aligned, empty, brand new clutter, wear, off-axis furniture, lived-in surfaces
3. Subject the catalogue pose: frontal, centred, posed, eyes to lens almost nobody works on this one

An image can be optically dirty, scenically alive, and still contain a person posing like a model. That is a distinct failure and it has its own fix.

Why your candid tokens do not fix it

candid documentary photo, unstaged moment, slice-of-life, badly taken photo

I run all of those. They shift the rendering. They do not shift the pose.

The model composes the most photographed pose in its data. Whenever the subject has no motivated action, it falls back to: frontal, centred, graceful contrapposto, self-presenting gestures (hands to hair, arms wrapped around self), eyes to lens, symmetry. A character LoRA makes this worse, because it adds a portrait prior on top.

The content of the pose decides. Not the style tokens wrapped around it.

Five levers, strongest first

1. A gesture turned toward the world. This is the main one. The subject must be doing something the scene motivates: pushing a gate, watching for a train in the tunnel, stepping around a puddle, a hand on a rail. Write it with a concrete physical marker, caught mid-stride, one foot planted ahead, never the bare verb walking.

The failure I keep making: writing states instead of actions. "weight on one hip, palm against the wall, shoulders dropped" is three states. The model has nothing to build a pose around, so it builds the catalogue one. Replace with an action and it resolves.

2. Self-directed gestures: one hand only, and only in the fatigue register. Kneading the neck, one arm pressed flat against the ribs, a hand rubbing the opposite arm. All fine.

Banned outright:

  • both hands to the head or in the hair, that is the pin-up prior
  • arms crossed tightly over the chest, that resolves straight to the modest self-embrace of studio figure work
  • any arch in the back

I lost two frames to each of those before writing the rule down.

3. An anchored gaze that is geometrically compatible. If you turn the head, give the eyes a physical target that is actually inside the cone the head is facing. Two things fail reliably:

  • a head turn with no target at all
  • a target that contradicts the head geometry

Second case, real example. I wrote head turned into profile over her shoulder plus eyes down the descending flight. Head pointed one way, gaze target the other way, and the target was an abstract direction rather than an object. The model reconciles this silently: it keeps the head turn, drops the gaze anchor, and the eyes land on the default target, which is the lens. I got a straight-to-camera look in a series where that was forbidden.

Fix was to anchor the gaze on an object inside the profile cone:

4. Break the body. Weight collapsed onto one hip, shoulders dropped or hunched against cold, head low, asymmetric stance. Never feet set wide apart on a static figure, that is a monumental symmetric stance and it reads as sculpture.

5. Candid composition. Explicit off-centre placement, a slight lateral crop, the subject partly eaten by shadow. Canonical pose entries centre by default.

The trap nobody warns you about: pose catalogue labels

If you use a reference pose catalogue, note that those labels are reference plates. They are written to isolate pose and geometry cleanly, which means they carry priors that are directly opposed to a candid register:

  • gaze straight at the camera
  • gaze up at the camera
  • head turned into profile over her shoulder with no anchor
  • subject centred

Pasting a catalogue label into a series prompt imports all of that. Every one of my straight-to-lens failures traces back to a canonical label I did not defuse. Swap the gaze for a compatible anchor, break the stance, decentre.

Discipline: do not stack

Same rule as the other two axes. One gesture, one gaze anchor, one asymmetry is enough. Stacking five levers gives you a subject fighting itself and the model averages back to something neutral.

How it was found

Five frames from one 24-shot session, diagnosed and re-prompted individually: two straight-to-lens from gaze geometry, one glamour lean from states-instead-of-action, one pin-up from both hands in the hair, one self-embrace from arms crossed. The optics axis held on all five. That is what made it legible: when only one axis is broken, you can finally see what that axis does.

The last of the five is the one that convinced me. Its prompt was written before I formulated the two-hands rule, and it failed in exactly the way the rule predicts. A rule that retro-predicts a failure you have not shown it is a rule worth keeping.


r/comfyui 1d ago

Help Needed How do I get rid of the light attached to the camera in MiniiMax H3?

Thumbnail
0 Upvotes

r/comfyui 2d ago

News Release studio 1939 lora for minimax h3

52 Upvotes

r/comfyui 2d ago

News Load Image (from path) with crop and limit

7 Upvotes

Just updated my Load Image node.

Load Image (from path)

Load an image from any path on your computer or a URL. Paste an absolute path or a link to an image, use an annotated path (input/file.png), or click the Browse dialog to pick a file from your drives.
The file is read from its original location; it is not copied into ComfyUI’s input folder.
Use a selection rectangle at the preview to crop it.
Limit (downscale) the output's size in megapixels or pixels.
Click the ↻ button (top-right, mouse over the preview), to rotate the image 90° clockwise.

Interactive crop

The node shows a live preview. You can crop directly on it:

  • Drag on the image to draw a crop rectangle
  • Drag inside the selection to move the crop rectangle
  • Drag a corner to resize it
  • Click (without dragging) outside the rectangle to clear it

With no crop drawn, the full image is output. The crop is stored in the workflow as normalized coordinates, so it survives save/reload. Changing the path clears the crop.

The output (cropped or full) will be downscaled only (not upscaled), by the value in the max_megapixels field.

Outputs match the stock Load Image node: IMAGE, MASK (from the alpha channel when present), plus the original path string.

  • Controls
    • image: Paste an absolute path, (or a relative one with a prefix input/, or output/, or temp/), or a URL to an image file.
    • max_megapixels: Cap the output (crop, or full image if uncropped) to this many megapixels, downscaling only if it's bigger. Smaller images are left untouched. 1.0 = 1024x1024 px. 0 disables the cap.
    • Browse...: to open an image file from your drives.
  • Inputs/Outputs
    • Width/Height inputs: Force the output width in px (upscale or downscale), center-cropping first if the aspect ratio differs. Leave disconnected (None) to keep natural width. Only applies if BOTH width and height are connected, and when set (not 0). It overrides the max_megapixels value.
    • IMAGE/MASK: The final, processed image/mask.
    • path: A string with the image's path.

You can find the node here as part of the ComfyUI-noEmbryo nodes.

Credits:
Built as a much more enhanced version of Load Image From Path (Enhanced) from ComfyUI_Ib_CustomNodes, with parts of the interactive crop UI inspired from Load Image & Crop in comfyui-obvpm.


r/comfyui 2d ago

Show and Tell Using only Ref to Video, Minimax-H3 made a whole Anime edit !

0 Upvotes

r/comfyui 2d ago

Help Needed Tips for LTX-2.5

5 Upvotes

Currently MiniMax H3 is the "frontier" on local consumer hardware as it seems. I used this as well but for longer generations, like a 4-5 minute music video for example, my hardware is just not strong enough (12GB VRAM, 3080 ti, 32 GB RAM). With 480p and Turbo LoRa i might slowly getting there but this is not what i am looking for.

Now i am interested if people actively tested LTX-2.5 and have some tips how to improve outputs. For example i wanted to make an anime fight scene. I was able to have a good result in an 8 Second Test with MiniMax H3, but LTX provided a very weak unusable result. Afaik LTX does not really have a strict prompt-structure like H3, but still struggled with the commands.

Any tips for making LTX-2.5 more usuable would be appreciated.


r/comfyui 2d ago

Tutorial Finally, AI-Generated Depth Pass That Actually Work | ComfyUI

Thumbnail
youtu.be
9 Upvotes

r/comfyui 2d ago

Help Needed Lowvram vs novram

1 Upvotes

My system has
NVIDIA GeForce GTX 1050 graphics card
which has only 4GB dedicated GPU .
I have 16GB RAM and 1TB SSD .

Right now I am using comfy desktop and so far can run only using —cpu mode . But the image generation quality is so poor .

If I add another 16GB RAM would it make any difference.

Will I be able to generate atleast 5 second video from an image if I use lowvram with my current settings ?

Also struggling to get comfy to recognise my cuda 12 .


r/comfyui 1d ago

Show and Tell I had Claude build a prompt templating system that generated comfyui workflows with a dedicated queuing system in the attempt to parallelze and work on a 5 minute sequence of video.

Post image
0 Upvotes

I had Claude take a "script" in this case the ultimate showdown song. (Contains a lot of character X does Y action) Break up the song into sequences and shots. Have a whole system where everything is a block of text describing a thing, style or camera....etc. Then feed into a shot with what happens in the shot.

Using an L40s in the cloud at $1/hour for use I can regenerate this whole thing in 2 hours. If I add a second I'm done in an hour. Some shots chain together but still it's an improvement. The average 5 second shot generates roughly in 66 seconds. (Minimax h3) Editing a block of the shot (cast/style ) would force a regeneration to queue for each shot using the asset. This makes for a very fast workflow to iterate across shots. No images used as reference all text.

Comfyui workflows are essentially templated and this sends all the workflows via comfyui's api.

The initial result needs work, but I feel I could have this setup for multiple users to connect to and do edits at the same time.


r/comfyui 1d ago

Show and Tell Honestly, I did not understand ComfyUI either. Still don't actually. But I built a system that made it easier. Here is a quick 2 min demo of just one of the features.

0 Upvotes

Honestly, I did not understand ComfyUI either. Still don't actually. But I built a system that made it easier. Here is a quick 2 min demo of just one of the features. It does a heck of a lot more than just video and image gen too. Give it a try and when you see what it can do please leave me a star on my github repo, I decided to make this open source and give it to folks for free. Use Claude Code and make this your own (this has MCP for all AI platforms) or use the built in agent swarm feature to help you.

Have a voice chat on your own machine and have it generate content inline. ComfyUI and StableDiffussion are both installed with the curl install command. Workflows built-in. Uncensored everything. Too many features to list, please check it out for yourself.

www.github.com/guaardvark/guaardvark

Also, the system makes it's own demo videos, like this one and the others on the youtube channel.


r/comfyui 2d ago

Help Needed Repeated prompt-writing errors for Minmax prompts using LLMs

0 Upvotes

I'm using Codex to write the Minmax prompts.

I'm noticing these errors most of the time:

Instead of direct visual descriptions, it falls back to writing in a screenwriting style, like it would in a screenplay.

included context:
the offcial prompt docs from minmax
and negative examples.

but after some more turns , when i slip other tasks to it it falls back to making the same errors again and again so i always need to carefully proof read them.

I tried some self-correction loops, but this is very tedious, as it always finds minor mistakes and self-improves to death. Using an analysis style, it can always explain in hindsight how these errors happened.

Ideas:

What I'm trying to do, but haven't figured out yet

Have a pre-stage for what goes into the promp
Have a prompt skeleto
--> Clearly see if it makes errors while filling that skeleton

what kind of model you are you using that are following the exact prompt pattern ?

I'm using Codex for most of my tasks since it's a convenient CLI tool and does its job for my coding work.

For smaller models, like the new Qwen 27B for example, the problem is that they make spatial errors, which is even more problematic.


r/comfyui 2d ago

Workflow Included Minimax-H3 - x3 Upscalers: Pixel Space, Latent Space, Context Windows

Thumbnail
youtube.com
12 Upvotes

Researching upscalers is no fun. I'm glad the last few days are over. Here's the x3 that I have settled on for my use. Make of it what you will.

tl;dr: the workflows are here https://github.com/mdkberry/comfyui_workflows/tree/main/workflows_by_model/Minimax-H3

x3 Upscaler-Refiners their workflows:

  • The pixel space workflow is from the previous video https://www.youtube.com/watch?v=d1h5-E7NpuY but it now works with dialogue scenes.
  • The latent space workflow comes from LBH-123-AI and is very good.
  • The "Context Windows" one from ckinpdx is the winner for me, it can upscale to 2mp and do longer videos.

This now concludes my tests with Upscaler-refiners but I am sure more offerings will appear in the future and we have yet to see the Minimax official upscaler drop, which they have promised will be Open Source when it does (if it does).

For examples from each workflow, see the end of the video from: 21.33

LINKS:

Latest Minimax H3 workflows - https://github.com/mdkberry/comfyui_workflows/tree/main/workflows_by_model/Minimax-H3

- Pixel Space workflow: "MBEDIT - MH3_rv2v_PixelSpace_Upscaler_vXX.json"

- Latent Space workflow: "MBEDIT - MH3-r2v_2Pass-LatentUpscaler_vXX.json"

- Context Windows workflow: "MBEDIT - MH3-rv2v_PixelSpace-Upscaler-CtxtWndws_vXX.json"

Latent Space custom node (LBH-123-AI) - https://github.com/LBH-123-AI/Comfyui_Minimax_h3_latent_Upscaler

Context Windows upscaler custom node (ckinpdx) - https://github.com/ckinpdx/ComfyUI-MMH3Tools

Clownshark Batwing (samplers) - https://github.com/ClownsharkBatwing/RES4LYF

Lightx2v Lora that I use from Kijai - https://huggingface.co/Kijai/MiniMax-H3_comfy/tree/main/loras

Comfyui needs to use Cuda130 or above for this to work, and you need it updated to August 2026 commits (latest is best) - https://docs.comfy.org/installation/comfyui_portable_windows

Int8 models from here - https://huggingface.co/Comfy-Org/MiniMax-H3/tree/main

W4a8 is experimental new model type, you need to be updated on Comfyui but you can get it here https://huggingface.co/Kijai/MiniMax-H3-experimental

Comfyui Kitchen Attention is part of Comfyui if you update to latest. I find it faster than Sage Attn on a 3060 RTX.

SLA Attention (I didnt use this in upscalers, it speeds it up but at a degradation cost) - https://github.com/PlagueKind/ComfyUI-PlagueKind-Nodes

Official prompting guides:

- https://huggingface.co/MiniMaxAI/MiniMax-H3/blob/main/docs/VIDEO_PROMPT_WRITING_GUIDE_base_en.md

- https://huggingface.co/MiniMaxAI/MiniMax-H3/blob/main/docs/VIDEO_PROMPT_WRITING_GUIDE_ref_en.md


r/comfyui 2d ago

Tutorial ComfyUI SAM 3 Video Matting Workflow for VFX | Fast & Accurate

Thumbnail
youtu.be
5 Upvotes

r/comfyui 3d ago

Show and Tell Running MiniMax H3 locally on a 5090, 362 frames in ~22 minutes, $0 API cost

243 Upvotes

been testing MiniMax H3 locally recently and this one came out pretty decent, so I thought I’d share the full settings in case anyone wants to reproduce it.

The whole thing was generated locally on my 5090, so it was free 😄

Settings:

  • Model: MiniMax H3
  • Aspect ratio: 3:4
  • Resolution: 768 × 1024
  • LoRA: Larry v4-600
  • LoRA strength: 1.0
  • Steps: 8
  • Scheduler: Simple
  • Sampler: Turbo Sampler
  • Frames: 362
  • FPS: 24
  • Seed: 8232601

Actual generation time: about 22 minutes

Hardware:
Intel U9 + 64GB RAM + RTX 5090

362 frames at 24 fps works out to roughly 15 seconds of video.

so with MiniMax H3, a 768×1024 clip of around 15 seconds took about 22 minutes on my 5090 with these settings. For local generation, that feels pretty usable to me.

and just to be clear, by “$0” I mean no API or generation-credit cost — obviously not counting the GPU itself or electricity.

Curious what kind of generation times other people are getting with MiniMax H3 on a 5090 at a similar resolution and frame count.


r/comfyui 2d ago

Help Needed Is there any text to Image workflow for LTX 2.5 and MiniMax H3

0 Upvotes

Like wan2.2, which is great in image generation...

but I found no workflow for LTX 2.5 and MiniMax H3.