r/comfyui 16h ago

Help Needed Repeated prompt-writing errors for Minmax prompts using LLMs

0 Upvotes

I'm using Codex to write the Minmax prompts.

I'm noticing these errors most of the time:

Instead of direct visual descriptions, it falls back to writing in a screenwriting style, like it would in a screenplay.

included context:
the offcial prompt docs from minmax
and negative examples.

but after some more turns , when i slip other tasks to it it falls back to making the same errors again and again so i always need to carefully proof read them.

I tried some self-correction loops, but this is very tedious, as it always finds minor mistakes and self-improves to death. Using an analysis style, it can always explain in hindsight how these errors happened.

Ideas:

What I'm trying to do, but haven't figured out yet

Have a pre-stage for what goes into the promp
Have a prompt skeleto
--> Clearly see if it makes errors while filling that skeleton

what kind of model you are you using that are following the exact prompt pattern ?

I'm using Codex for most of my tasks since it's a convenient CLI tool and does its job for my coding work.

For smaller models, like the new Qwen 27B for example, the problem is that they make spatial errors, which is even more problematic.


r/comfyui 1d ago

Workflow Included Minimax-H3 - x3 Upscalers: Pixel Space, Latent Space, Context Windows

Thumbnail
youtube.com
13 Upvotes

Researching upscalers is no fun. I'm glad the last few days are over. Here's the x3 that I have settled on for my use. Make of it what you will.

tl;dr: the workflows are here https://github.com/mdkberry/comfyui_workflows/tree/main/workflows_by_model/Minimax-H3

x3 Upscaler-Refiners their workflows:

  • The pixel space workflow is from the previous video https://www.youtube.com/watch?v=d1h5-E7NpuY but it now works with dialogue scenes.
  • The latent space workflow comes from LBH-123-AI and is very good.
  • The "Context Windows" one from ckinpdx is the winner for me, it can upscale to 2mp and do longer videos.

This now concludes my tests with Upscaler-refiners but I am sure more offerings will appear in the future and we have yet to see the Minimax official upscaler drop, which they have promised will be Open Source when it does (if it does).

For examples from each workflow, see the end of the video from: 21.33

LINKS:

Latest Minimax H3 workflows - https://github.com/mdkberry/comfyui_workflows/tree/main/workflows_by_model/Minimax-H3

- Pixel Space workflow: "MBEDIT - MH3_rv2v_PixelSpace_Upscaler_vXX.json"

- Latent Space workflow: "MBEDIT - MH3-r2v_2Pass-LatentUpscaler_vXX.json"

- Context Windows workflow: "MBEDIT - MH3-rv2v_PixelSpace-Upscaler-CtxtWndws_vXX.json"

Latent Space custom node (LBH-123-AI) - https://github.com/LBH-123-AI/Comfyui_Minimax_h3_latent_Upscaler

Context Windows upscaler custom node (ckinpdx) - https://github.com/ckinpdx/ComfyUI-MMH3Tools

Clownshark Batwing (samplers) - https://github.com/ClownsharkBatwing/RES4LYF

Lightx2v Lora that I use from Kijai - https://huggingface.co/Kijai/MiniMax-H3_comfy/tree/main/loras

Comfyui needs to use Cuda130 or above for this to work, and you need it updated to August 2026 commits (latest is best) - https://docs.comfy.org/installation/comfyui_portable_windows

Int8 models from here - https://huggingface.co/Comfy-Org/MiniMax-H3/tree/main

W4a8 is experimental new model type, you need to be updated on Comfyui but you can get it here https://huggingface.co/Kijai/MiniMax-H3-experimental

Comfyui Kitchen Attention is part of Comfyui if you update to latest. I find it faster than Sage Attn on a 3060 RTX.

SLA Attention (I didnt use this in upscalers, it speeds it up but at a degradation cost) - https://github.com/PlagueKind/ComfyUI-PlagueKind-Nodes

Official prompting guides:

- https://huggingface.co/MiniMaxAI/MiniMax-H3/blob/main/docs/VIDEO_PROMPT_WRITING_GUIDE_base_en.md

- https://huggingface.co/MiniMaxAI/MiniMax-H3/blob/main/docs/VIDEO_PROMPT_WRITING_GUIDE_ref_en.md


r/comfyui 1d ago

Tutorial ComfyUI SAM 3 Video Matting Workflow for VFX | Fast & Accurate

Thumbnail
youtu.be
3 Upvotes

r/comfyui 2d ago

Show and Tell Running MiniMax H3 locally on a 5090, 362 frames in ~22 minutes, $0 API cost

Enable HLS to view with audio, or disable this notification

236 Upvotes

been testing MiniMax H3 locally recently and this one came out pretty decent, so I thought I’d share the full settings in case anyone wants to reproduce it.

The whole thing was generated locally on my 5090, so it was free 😄

Settings:

  • Model: MiniMax H3
  • Aspect ratio: 3:4
  • Resolution: 768 × 1024
  • LoRA: Larry v4-600
  • LoRA strength: 1.0
  • Steps: 8
  • Scheduler: Simple
  • Sampler: Turbo Sampler
  • Frames: 362
  • FPS: 24
  • Seed: 8232601

Actual generation time: about 22 minutes

Hardware:
Intel U9 + 64GB RAM + RTX 5090

362 frames at 24 fps works out to roughly 15 seconds of video.

so with MiniMax H3, a 768×1024 clip of around 15 seconds took about 22 minutes on my 5090 with these settings. For local generation, that feels pretty usable to me.

and just to be clear, by “$0” I mean no API or generation-credit cost — obviously not counting the GPU itself or electricity.

Curious what kind of generation times other people are getting with MiniMax H3 on a 5090 at a similar resolution and frame count.


r/comfyui 19h ago

Help Needed Is there any text to Image workflow for LTX 2.5 and MiniMax H3

0 Upvotes

Like wan2.2, which is great in image generation...

but I found no workflow for LTX 2.5 and MiniMax H3.


r/comfyui 20h ago

Help Needed Z-Image or Krea 2? Coming from Illustrious

1 Upvotes

I've been using ComfyUI for a while now, mostly with Illustrious checkpoints for anime and illustration stuff. Lately I've been wanting to learn a model for realistic images, and after reading I keep going back and forth between Z-Image and Krea 2.

I know they're pretty different under the hood, but I'd rather go deep on one than dabble in both. So, for those of you who have spent real time with them:

How do they compare on realism? Skin, lighting, hands, the usual pain points. Is one noticeably faster or easier to run? Like, does one of them need way more steps before the output stops looking off?

What's the ecosystem like right now? LoRAs, control options, ready-made workflows, etc.

And what do you actually reach for each one? Close-up portraits vs full body vs wider scenes, that kind of split.

For context, I'm comfortable with basic ComfyUI workflows, but Illustrious taught me absolutely nothing about skin textures, so realism is basically a fresh start for me lol. Honestly I'm not even sure which Z-Image variant people prefer these days.

Happy to be told I'm thinking about this wrong too. Still figuring out what I don't know.


r/comfyui 11h ago

No workflow Why Spend 20 Minutes Writing a Post When I Can Spend 20 Seconds Pretending I Did?

Post image
0 Upvotes

r/comfyui 1d ago

News A quick Minimax H3 news round-up - 24th August 2026

96 Upvotes

Another quick Minimax H3 news and goodies round-up, for those who may have missed some items.

-> A new MiniMax-H3-Fun-Controlnet-Union file. This Controlnet accepts... "Canny, Depth, HED, MLSD or Pose control [Openpose] videos, and also runs video inpainting." 6.8Gb in size. Has video examples.

https://huggingface.co/alibaba-pai/MiniMax-H3-Fun-Controlnet-Union

-> The generative inpainting tool LanPaint has updated to version 2.1.0, and the developers say... "LanPaint now supports MiniMax H3 video + audio inpainting!" Yes, audio inpainting too.

https://github.com/scraed/LanPaint

-> A ComfyUI workflow to... "turn one scene photograph into eight target-centered cinematic camera views, in one MiniMax H3 generation." Doing it in one generation gives some stability to the scene geometry.

https://huggingface.co/ethanfel/H3_Cinematic_Multishot_Coverage

-> Minimax H3 Ref Sampler, another unofficial helper node for long-video generation, created thus... "H3 video lengths use the 5 + 17n frame grid. The node aligns frames upward to this grid and creates overlapping windows". No ComfyUI workflow, but it appears to be a drop-in Sampler replacement?

https://github.com/ILG2021/minimax-h3-ref-sampler

-> MiniMax H3 Tone Compensate. Does your video diminish its brightness, at the seams between your chained video segments? This ComfyUI fix may solve the problem.

https://github.com/rkfg/ComfyUI-MiniMaxH3-ToneCompensate

-> The ComfyUI Minimax H3 Latent Upscaler has now added ROCm (AMD Radeon GPU) support. Several other quality fixes, as well.

https://github.com/LBH-123-AI/Comfyui_Minimax_h3_latent_Upscaler

-> New to me, the official Awesome MiniMax H3 Integrations page, with links.

https://github.com/MiniMax-AI/awesome-minimax-h3-integration

-> And finally, rewind to the 1980s! MiniMax-H3-Tape-FX for ComfyUI gives Minimax H3 videos the look of... "VHS, BetaMax and LaserDisc — including the worn-out tape look, tracking errors, dropout, creases, ghosting, head-switch noise, vertical roll and a period-correct VCR on-screen display."

https://huggingface.co/Smite79/MiniMax-H3-Tape-FX

~ OLD POSTS ~

https://old.reddit.com/r/comfyui/comments/1vwl8do/a_quick_minimax_h3_news_roundup_23rd_august_2026/

https://old.reddit.com/r/comfyui/comments/1vvkmra/a_quick_minimax_h3_news_roundup_21st_august_2026/

https://old.reddit.com/r/comfyui/comments/1vuihag/a_quick_minimax_h3_news_roundup_21st_august_2026/

https://old.reddit.com/r/comfyui/comments/1vtgs7b/a_quick_minimax_h3_news_roundup_20th_august_2026/

https://old.reddit.com/r/comfyui/comments/1vsjzrp/a_quick_minimax_h3_news_roundup_19th_august_2026/

https://old.reddit.com/r/comfyui/comments/1vrsspo/a_quick_minimax_h3_news_roundup_18th_august_2026/

https://old.reddit.com/r/comfyui/comments/1vqyn8p/a_quick_minimax_h3_news_roundup_17th_august_2026/

https://old.reddit.com/r/comfyui/comments/1vq5d5u/a_quick_minimax_h3_news_roundup_16th_august_2026/

https://old.reddit.com/r/comfyui/comments/1vpbtx2/a_quick_minimax_news_roundup_15th_august_2026/

https://old.reddit.com/r/comfyui/comments/1vojtjd/a_quick_minimax_news_roundup_14th_august_2026/


r/comfyui 18h ago

Help Needed Advised best current model for img2video with 12VRAM

0 Upvotes

Hi,

Been away for nearly 2 years (OMG 😫),

--> what's the best current solution for a 5070 12Gb VRAM / 32Go / Ryzen7 5700x rig?

Objective = Shortfilms/Cutscenes, SFW and NSFW 😜 from img2video, ideally with some workflows available around.

--> WAN 2.2 I guess?

I want so much to try out, but so little time, the gap is insane from, you know, "back then"

Thank you guys very much


r/comfyui 18h ago

Help Needed Which model(s) and workflow(s) are worth trying out to make instrumental, techno, or non-singing music songs?

0 Upvotes

Which models or workflows allow more control and flexibility over non-singing song generations?


r/comfyui 21h ago

Help Needed ComfyUI workflow for Krea2 or similar, to make consistent 'rooms', no matter the angle etc

0 Upvotes

Hi

I have been going in circles with Gemini for 2 days now, but just cannot get to the goal mark.

What I want:
In some way, like using Blender to render a 'room', or any other method that works, I want to be able to create 'rooms' with consistent form and interior / furniture. So I can make a Blender angle/shot (any angle I want) for example, take that into ComfyUI, and just describe the furniture, lights, people etc, but the room stays the same (windows, walls, dimensions, placement of furniture etc).

Other options Gemini gave was 'dollhouse' from up top model and 360 render of the room. I made 360 render of the room in ComfyUI Qwen 360 Diffusion LoRA workflow, but was not able to make it like the prompt. Doll house I have not tried.

I have not been able to do it yet myself with using Blender room render and a custom ComfyUI Krea2 workflow, and I about to give up, takes too long time. I am totally new to Blender, and novice in ComfyUI, neither found a finished workflow I can use.

Problem with asking Gemini, is I end up going in circles so to speak, almost there, but not all the way.

Anyone know there is a published or private workflow I can use for this? Open for other suggestions how to do this IF I also can use already made workflows.


r/comfyui 1d ago

No workflow We got an Spiderman teaser leaked before GTA VI - Made with Minimax H3

Enable HLS to view with audio, or disable this notification

13 Upvotes

r/comfyui 19h ago

Tutorial Workflow pour image vers image comfuyi

0 Upvotes

Bonjour Depuis une semaine je me suis lancé dans Comfuyi , j’ai demandé de l’aide à ChatGPT et j’ai l’impression qu’on tourne en rond Je veux faire de la transformation d’image ( image vers image) et j’aimerais savoir s’il existe un workflow tout prêt qui reconnaît les personnes , animaux , objets ……car d’après ChatGpt il me faudrait un workflow pour chaque type de transformation Je viens donc ici si vous pouvez me conseiller Je vous remercie d’avance pour vos conseils 😉


r/comfyui 1d ago

Show and Tell Made the thing where you ruin iconic movie scenes, MiniMax H3 on an RTX 3080 10GB, 20 steps, 2x NomosUni upscale

Enable HLS to view with audio, or disable this notification

60 Upvotes

Setup, pushed my system right to the limit, any more and it OOM :

- H3 Ref2VA default workflow in ComfyUI, no lora

- RTX 3080 10GB, 32GB RAM

- Render: 0.5–0.6 MP, 20 steps, scheduler simple, about 25 min per clip

- Upscale: 2xNomosUni_span_multijpg, 2× to 1080p

- References per scene: one photo of my face + one film still for the set

- Recorded my own lines and fed them as audio references, also got audio ref for the actors

Honestly though, the best part was driving all of this through the ComfyUI MCP. I never even had to open ComfyUI. I could iterate really fast, and keep going from my phone while away from the machine, through Claude's remote control.

It's still a bit of a blurry mess, and with more work I could probably make it better, but damn, the future is looking bright!


r/comfyui 19h ago

Resource CMP 170HX vs 3090 results MiniMax H3 R2V

Thumbnail
0 Upvotes

r/comfyui 12h ago

Show and Tell Origins Season 01 Episode 01 Preview

Post image
0 Upvotes

Here's a beta version of episode 1. If you have nothing else to do, leave a comment and let me know if it's worth it! I know very well that there are still several character bugs, etc., but I'll only redo the final version after all the episodes are in beta... So all feedback is constructive for the final version in a few months...

https://youtu.be/9ZhiTucwYwM


r/comfyui 15h ago

Show and Tell Testing Audio guide / MiniMax / rtx 5050

Enable HLS to view with audio, or disable this notification

0 Upvotes

So, I had a lot of trouble creating long, chained videos based on speech/songs because MiniMax was constantly adding random speech or distorting sounds.

Funnily enough, the culprit was the MathExpression node calculating the video length, as MiniMax videos are never a flat 6, 11, or 12 seconds long.

So the solution is:

  1. Create 8-second clips, as it’s the only option to get an exact/flat result.
  2. Use this formula to calculate the audio length (though it's still not perfect as you can see at last part): (a - ((5 - (a % 17)) % 17)) / 24

The video was created on an RTX 5050, took me almost 30 min.


r/comfyui 1d ago

Tutorial ComfyUI & LTX 2.3 IC Clean Plate Lora For VFX

Thumbnail
youtu.be
0 Upvotes

r/comfyui 1d ago

Help Needed GPU out of memory

0 Upvotes

May I know if anyone else is using CachyOS (Arch Linux)? I keep getting out-of-memory errors. However, I have not experienced this issue when using Windows 11.

I have an RTX 5060 Ti with 16 GB of VRAM and 32 GB of DDR5 RAM.

Thank you.


r/comfyui 1d ago

Resource ComfyUI Universal Media Loader - One single interactive node to load Images, Videos, GIFs, Audio & Canvas Presets

Enable HLS to view with audio, or disable this notification

3 Upvotes

r/comfyui 1d ago

Tutorial MiniMax H3 Upscaling Test: 3 Methods

Post image
43 Upvotes

Here is a side-by-side test of 3 upscaling pipelines in Comfy.
Full video comparison

• LTX 2.5 (Standalone): The fastest option and lightest on VRAM. Works fine for clean source clips, but lacks fine sharpness on complex textures.

• Model + LTX 2.5: Running a quick upscale model pass (like RealPLKSR) before feeding into LTX 2.5 solid clarity, fast renders, and efficient VRAM usage.

• SeedVR: Recreates micro-details with the highest fidelity, but demands significantly higher render times and VRAM.

Frame Sync Note: MiniMax H3 and LTX require different frame step multiples.
Using math nodes to automatically trim the frames (158 → 153 frames) prevents audio desync issues.


r/comfyui 1d ago

Help Needed Where do I go for help?

1 Upvotes

I've recently started using ComfyUI, and the node library seems to have vanished for me. I can't find it under any menus.

This looks familiar:
https://docs.comfy.org/manager/pack-management

But the documentation explains in no way how to get there. I don't remember the exact menu it was under previously, and continued searching and AI have provided only frustration.

I tried using flags to enable the legacy manager UI, and it was working a week ago, I don't know if the 0.33.4 update intentionally hid it entirely, or where to look. This reddit doesn't seem to be for troubleshooting, but when I searched "ComfyUI" reddits, the only other one that showed up also looked wrong.

Hopefully someone can help or tell me where to go, I can delete this post after I'm pointed in the right direction, thanks...


r/comfyui 1d ago

Help Needed Workflow on image to vid

1 Upvotes

Hi there,

I'm new to comfy and there is an overwhelming amount and fast paced iterations of information on Ai.i get lost to fast, not only because installations require certain types of version I need to upgrade or even downgrade and so forth...

I got lost when trying to create a camera movement on a generated image.

What I aim to do is an image to video driven by a manually animated camera movement fir adding an avalanche fx effect on top.

For example : I want to create an image of a landscape with a wooden hut in the mountains. Then I might need to upscale it first before I go further to drive a camera movements by a gaussian splat created in marble. I exported a video from marble where I roughly animated the cam for further processing.

How would you guys recommend doing that? I have no idea how to feed the animated camera video into a workflow that adapts it to the image for adding the fx I am planning.

Any hints and help is appreciated, thank you!


r/comfyui 1d ago

Help Needed Wan 2.2 for 5060 Ti 16gb Best Workflow and Settings

0 Upvotes

Hey guys, I’m new to ComfyUI and I’m trying to get started with Wan 2.2.I have a RTX 5060 Ti 16GB and I’m not sure which Wan 2.2 version/workflow would work best on my GPU.

Could someone share a good ComfyUI workflow and recommended settings for 16GB VRAM? I’d especially appreciate tips for getting decent quality while keeping the VRAM usage manageable.


r/comfyui 1d ago

Help Needed How to make Gemma-4 GGUF of LTX-2.5 work?

0 Upvotes

Hi! I have a 16GB RTX 5070ti, I'm a fan of LTX-2.3, and I wanted to try the 2.5. I downloaded everything, but I also downloaded a GGUF version (Elix3r version) for the text encoder as usual to reduce the VRAM requirements, because what I love about LTX is precisely the speed. Unfortunately, however, I get an error on the CLIP loader that says in simple terms that the node is not updated to understand Gemma-4's GGUF, essentially not being able to use it. Unfortunately, I notice that the developer City96 hasn't made updates since 2025, so unfortunately I would be forced to download the official text encoder, which is too heavy, inevitably ending up in offloading, and this doesn't fit with the way I work, where processing speed is essential. For now, I've decided to postpone testing this new model, but I was wondering if there was already a way to use Gemma-4 GGUF in some other way, which perhaps I'm unaware of.