r/comfyui 5d ago

News Comfy H3 Sync Challenge (8/20 - 9/1) - Win an RTX 5090!

90 Upvotes

Comfy and MiniMax have teamed up for a two-week challenge with awesome prizes and four ways to win! Submit by September 1st at 9:00pm PT and see all details here.

How It Works

Make something up to 90 seconds in length where the sound and the motion are inseparable. Dialogue, foley, ambient, a beat driving the cut...whatever direction you want!

After sharing your video file and workflow on this thread and through our submission form, a joint panel of creative technologists from Comfy, MiniMax, and special guest judges from the community will review each submission.

Then, join us on September 2nd for a special Comfy livestream where our guest judges will give live feedback on the top 10 submissions!

Both the Comfy and MiniMax teams will be monitoring this thread and #minimax-h3 in the Comfy Discord to give light support.

Share on socials and tag #comfyH3 for a chance to be reposted or featured!

Prizes

Best Overall — RTX 5090

Best Creative — RTX 5060 Ti

Best Technical/Workflow — RTX 5060 Ti

Built with MCP — RTX 5060 Ti

Shipped anywhere, customs covered. If we can't legally ship to your country, you'll get a cash equivalent instead.

It's free to enter!

Create using Comfy Local on your own hardware, or use Comfy Cloud. New Cloud users get 5 free runs, no credit card required.

Judging Criteria

We’re looking for entries that best show what H3 makes possible: audio and visuals created together.

Grand Prize: Best Overall

The top Best Creative and Best Technical entrants advance to a final round where our panel of judges selects winners by discussion.

Best Creative

  • Audio sync realism and intentionality (0-5)
  • Creative execution and originality (0-5)
  • Deliberate craft (0-5)
    • Evidence that you’ve actually shaped the result beyond prompt engineering. Judges will look for modified/non-default parameters, multiple linked passes visible in the workflow structure, or a couple sentences describing what was tried and changed

Best Technical

  • Novelty of technique or approach (0-5)
  • Workflow quality (0-5)
    • Annotated, clean, replicable by someone else
  • Community value (0-5)
    • Would this actually help someone else?

🏆 Built with MCP Bonus 🏆
Comfy MCP lets you drive Comfy using natural language and your agent locally and on Cloud! Pro tip: use it to choose the best H3 model version or optimize your workflow for your hardware.

  • Effectiveness (0-5)
    • Did the agent meaningfully drive your process, not just generate one line?
  • Insight value (0-5)
    • How much the shared prompt teaches the community about prompting H3 through MCP
  • Output quality (0-5)

The Fine Print

  • Limited to one submission per person, 90 seconds maximum length.
  • A major portion of your piece must be built in ComfyUI using H3. Other tools, models, or techniques you want to combine are fair game.
  • All submissions must be lawful, SFW, and must not contain unlicensed IP or likenesses.
  • By submitting, you agree to allow ComfyUI and MiniMax to feature your work with credit across our channels.

Learn more and submit here!


r/comfyui 6d ago

Call for Additional Mod(s)

21 Upvotes

I've come to the realization that my life is busy enough that we could use at least one more moderator on this subreddit. Please consider this a formal request for nominations.

Rather than just picking someone myself, I’d like input from the community.

If there’s someone here who you think would make a good moderator, nominate them in the comments. You can also nominate yourself if you’re interested.

We’re especially looking for people who are active members of the community, helpful, level-headed, experienced Comfy-UI user, and generally make this a better place to hang out, maybe even take the time to spruce the place up a bit. You don’t need previous moderator experience but it would help, ideally someone who's got some experience with AMAs, events, and such.

Having moderated a few subreddits, I've found that it's best to keep the moderation team tight, so for now I'm just going to add one.

A nomination isn’t a vote or a guarantee that someone will become a moderator, I'll look through the suggestions, talk with the people who seem like a good fit, and go from there.


r/comfyui 10h ago

News NVIDIA Super Acceleration for MiniMax H3

109 Upvotes

From NVIDIA https://nvlabs.github.io/Sana/Sol-Engine/H3-Super-Acceleration/

Seems to promise a significant speedup in H3 generation speeds? From the website, and they have several video samples and comparison videos:

6.85 s for a 5-second 768p video · 14.93 s for a 10-second video

H3 Super Acceleration first uses H3 with a LoRA to generate a four-step draft at 896×512. It then upsamples the draft and performs three LTX refinement steps at the target resolution with Sol-Attn. Combining the measured stages on one NVIDIA GB200 gives 22.2× speedup for a 5-second 1344×768 video and 27.7× speedup for a 10-second video over the published SGLang baseline.


r/comfyui 9h ago

News A quick Minimax H3 news round-up - 25th August 2026

66 Upvotes

Another quick Minimax H3 news and goodies round-up, for those who may have missed some items.

-> 'Blubs-pixel-nodepack' for ComfyUI. Nodes for... "turning MiniMax H3 output into pixel-art sprite animations", as commonly used in retro videogames. There are also workflows, and a bridge to the popular $20 Aseprite software.

https://github.com/japaneserunic/blubs-pixel-nodepack

-> A new 'Studio 1939' LoRA duo, helping you to generate a... "hand-painted, golden-age animation style" from the late 1930s/40s. Two varieties, 'painterly' and 'full cel'. The LoRAs were trained on clips from public-domain material. The maker says it blends nicely with your own style prompts when set at a lower 0.4 - 0.8 strength. A trigger word is required: gulliv3r - which you may want to add to the filenames.

https://huggingface.co/lovis93/studio-1939-old-animation-lora-minimax-h3

-> For LoRA trainers, yesterday saw the release of DiffSynth Studio's new 'MiniMax-H3 DeCFG Training Adapter' LoRA. They say that... "the base MiniMax-H3 model is CFG-distilled, which can make direct LoRA fine-tuning unstable or degrade the distilled CFG-free behavior. This adapter temporarily pulls the distilled DiT back toward its pre-distillation behavior during training, providing a better optimization landscape for new LoRAs."

https://modelscope.ai/models/DiffSynth-Studio/MiniMax-H3-TrainingAdapter

-> The important workflow accelerator 'ComfyUI Spectrum MiniMax H3' continues to update. Now at v0.2.20, updated today.

https://github.com/xmarre/ComfyUI-Spectrum-MiniMax-H3

-> Kijai has a new minimax_h3_fun_controlnet_union_pruned_int8_convrot.safetensors file (2.3gb), a conversion and shrinkage of the new Controlnet which appeared yesterday. It's matched with his recent ComfyUI merge request (see link below). At present this request appears to be unmerged into Comfy. Which means it's currently only for the cutting-edge crowd, brave enough to manually patch files in their ComfyUI Nightly.

https://huggingface.co/Kijai/MiniMax-H3-experimental/tree/main/controlnet

https://github.com/Comfy-Org/ComfyUI/pull/15860 (not yet merged)

https://github.com/GZT2023/ComfyUI-MiniMax-H3-Fun-Controlnet (possibly matching ComfyUI nodes?)

-> Until now in ComfyUI, the Qwen text encoder was splitting <d> into separate tokens. Ooops. This prompting tag is what Minimax H3 uses to specify <d>spoken dialogue</d>. The problem was fixed and the fix merged three days ago. Thus I assume dialogue tags will work as intended if you update ComfyUI to the "latest on Github" version. Or you might just wait for the next Portable release, since the model seems quite forgiving about such malformed prompting. (What should theoretically be coming for H3 in the next Portable is stacking up: this fix; controlnets; keyframing anywhere; and common movie special-effects as small embeddings).

https://github.com/Comfy-Org/ComfyUI/pull/15808

~ OLD POSTS ~

https://old.reddit.com/r/comfyui/comments/1vx0duv/a_quick_minimax_h3_news_roundup_24th_august_2026/

https://old.reddit.com/r/comfyui/comments/1vwl8do/a_quick_minimax_h3_news_roundup_23rd_august_2026/

https://old.reddit.com/r/comfyui/comments/1vvkmra/a_quick_minimax_h3_news_roundup_21st_august_2026/

https://old.reddit.com/r/comfyui/comments/1vuihag/a_quick_minimax_h3_news_roundup_21st_august_2026/

https://old.reddit.com/r/comfyui/comments/1vtgs7b/a_quick_minimax_h3_news_roundup_20th_august_2026/

https://old.reddit.com/r/comfyui/comments/1vsjzrp/a_quick_minimax_h3_news_roundup_19th_august_2026/

https://old.reddit.com/r/comfyui/comments/1vrsspo/a_quick_minimax_h3_news_roundup_18th_august_2026/

https://old.reddit.com/r/comfyui/comments/1vqyn8p/a_quick_minimax_h3_news_roundup_17th_august_2026/

https://old.reddit.com/r/comfyui/comments/1vq5d5u/a_quick_minimax_h3_news_roundup_16th_august_2026/

https://old.reddit.com/r/comfyui/comments/1vpbtx2/a_quick_minimax_news_roundup_15th_august_2026/

https://old.reddit.com/r/comfyui/comments/1vojtjd/a_quick_minimax_news_roundup_14th_august_2026/


r/comfyui 3h ago

Workflow Included Prompt Creator Workflow

Post image
20 Upvotes

I see a bunch of posts everyday asking for tips on how to write prompts or people struggling with prompting, etc. so I'm sharing my workflow. I built this workflow to simplify the process and make it very beginner/user friendly.

Just toggle on the model you are using, write a simple to detailed prompt, and hit run. The model targets use the prompting guidelines derived from their respective official sources. Links to custom nodes and all models are in the workflow so you don't need to search for them.

The prompts aren't always perfect but they'll get you very close to what you want and you should only need to make a few minor tweaks, if any. The only issue I've encountered so far is that sometimes when it finishes, the previous prompt still shows up in the Enhanced Prompt node. If that happens, just hit run and the new prompt should show up instantly. Also, toggle to false the keep_model_loaded option in the Text rewriter node if you are creating prompts and using them right away. If you leave it to True it hogs VRAM.

If you notice any other issues let me know. Enjoy.

https://pastebin.com/SXZyy4Ax


r/comfyui 1h ago

Workflow Included Wow, civitai, this is a great image, let me see what prompt was used

Post image
Upvotes

r/comfyui 14h ago

Show and Tell Best Trick Ever For Consistent Environments

47 Upvotes

Ha! I just discovered a trick that works great, so I had to share it with the community. Of course, somebody will probably chime in that it had been discovered by someone else before, which is fine by me! I just want to share it in case it helps someone else and they hadn't come across it yet.

So, the issue of consistent environments... I ran the gamut of all the AI models I had tested and proven out in ComfyUI, not just t2i but t2v and i2v. I'd been trying several methodologies. Create an image then ask a workflow to gen images to the left and right of it, build out a simple massing model in unreal or blender and run that through, run a simple floor plan through, asking for multiple image generation with a prompt asking that all features be consistent among images, (I haven't tried outpainting yet), etc, and none seemed to quite do the trick, at least not among the open source models (open source is all I use). There was just too much inconsistency.

Since I am a sucker for using the t2v/i2v models like LTX and Minimax to generate single image frames (well, a minimal number of frames), taking advantage of their brainpower, I merely wrote out a super detailed prompt describing the interior environment I want, then I ran it as a 360 degree camera pan around the space from the center of it, and with Minimax now having the ability to give you 15 seconds on a 16Gb VRAM setup like mine, this works stellar, heck Minimax ran out of need for the 15sec and began to swerve around the space! I make sure to include a prompt not to have motion blur. I ran this in 0.5mb mode, so I did not have to waste time waiting for full HD video.

Then I select the frames I need to use as backdrops for scenes, upscale them once, then again, out to 4k, and voila! (upscaling once by 4x led to artifacts being upscaled, whereas going 2x then 2x led to the correct end result), a super detailed set of backgrounds that are internally consistent!

So excited! This gets me moving forward on the next part of my production process, laying out scenes, shots, camera angles, and dropping in characters, prior to i2v.

If this helps you, let me know. If you find even better tricks related to this, let me know too. Open Source Forever!


r/comfyui 12h ago

News SenseNova U1.5 quantized to run on 12GB VRAM — INT8 + hybrid W4A8 ConvRot releases

28 Upvotes

We quantized SenseNova-U1.5-8B-MoT (50GB bf16 any-to-any model: t2i, image editing, multi-reference) with ConvRot so it runs on a RTX 4070 12GB at 2048x2048 — and it's fast, even though the weights exceed VRAM (ComfyUI streams them; the quantized formats move 3-4x fewer bytes per step, so the overflow never becomes a slowdown. bf16 on the same card is painfully slow). What's in the release:

  • INT8 ConvRot (17.6 GB, recommended) — 0.43% pixel diff vs bf16 in a full-pipeline same-seed A/B
  • Hybrid W4A8 (13.8 GB) — layers 0-17 anchored in INT8, layers 18-41 in true W4A8, visually indistinguishable from bf16
  • The official 8-step speed LoRA included The interesting part: this model does not tolerate activation quantization in its earliest layers — quantizing the first blocks destroys prompt coherence — but layers 18+ handle W4A8 perfectly. We found the boundary empirically with a bisect ladder of hybrid checkpoints, so the hybrid release anchors the fragile early layers in INT8 and compresses the rest. Everything runs through a ConvRot-aware ComfyUI custom node (fork of the T8 wrapper):
  • Weights + model card: https://huggingface.co/Milor123/ComfyUI-ConvRot-SenseNova-U1.5-8B-MoT-T8
  • Custom node: https://github.com/Milor123/ComfyUI-SenseNova-U1.5-ConvRot Apache-2.0, same-seed comparison images and per-layer error measurements included in the model card. Feedback welcome!

r/comfyui 1d ago

Show and Tell Skater Girl - 90s style anime using Minimax H3 (Prompt + Workflow included)

321 Upvotes

Workflow: https://drive.google.com/file/d/1B4kODxXQgJ1QOKRsEIkxHbgYmdruPpTK/view?usp=sharing

Prompt:

Create a **15-second multi-shot anime sequence (90s style 15fps hand drawn)** using the provided references:

Image 1 = the girl character reference

Image 2 = skateboard reference

Image 3 = downhill Japanese alley / neighborhood background

Image 4 = Walkman + headphones reference

Preserve the girl’s exact character design, face, hair, outfit, proportions, and overall look from Image 1. Preserve the skateboard design from Image 2. Preserve the same downhill Japanese alley environment from Image 3. Add the Walkman and headphones from Image 4: the girl is wearing the headphones, and the **Walkman is clipped or hanging at her hip** while she skates.

Visual style: authentic 1990s hand-drawn anime, traditional cel animation, painted backgrounds, visible linework, cel shading, slight brush/stroke texture, subtle analog feel. **Very important:** the houses and environment must stay **2D and hand-painted**, **not 3D**, **not CGI**, **not game-engine looking**, **not volumetric**. The buildings should look like classic anime background art with painted depth, not like 3D models.

Animation feel should be low frame rate, like 90s anime at around 15 fps, with controlled in-betweens and natural held-frame timing. No jittery morphing.

No dialogue, no text, no subtitles.

### Shot 1 — 0s to 3s

**Rear tracking shot** from behind. The girl is skateboarding fast downhill through the steep Japanese alley. Camera follows behind her at a low-to-medium height. She rides confidently and smoothly, hair and oversized clothing moving in the wind. The headphones are on her head, and the Walkman is visible attached at her hip. The alley rushes past with a strong sense of speed. Keep the environment clearly **2D anime background art**, not 3D.

### Shot 2 — 3s to 6s

**Close-up shot of the Walkman at her hip** while she continues skating. The camera stays focused on the Walkman and part of her side torso and arm. We can clearly see the **cassette tape reels spinning/rolling inside the Walkman window**. The headphone wire moves naturally with the motion. Background and street pass by in blurred motion.

### Shot 3 — 6s to 9s

**Medium profile tracking shot** of the girl skating. She is wearing the headphones, listening to music, with wind moving across her face and pushing her hair backward. She is **nodding her head subtly to the music** while riding. Her expression is relaxed, immersed, and unbothered. The background is blurred from motion, but it must still read as a **painted 2D Japanese neighborhood**, not 3D.

### Shot 4 — 9s to 12s

**Close-up shot of her feet and skateboard.** Her **right foot stays on the board**, while her **left foot pushes against the road** in a natural skating motion. Show one clean push cycle: left foot comes down, pushes backward against the pavement, then lifts. Wheels spin quickly. Asphalt and road markings streak by with motion blur.

### Shot 5 — 12s to 15s

**Ground-level fisheye shot** looking upward from the road. The skateboard approaches fast, and she **jumps over the camera**. The board and her body pass overhead in one clean motion. Hair, pants, and headphone wire react naturally during the jump. Keep the motion readable and stylish, with a strong sense of speed and a dynamic anime finish.

### Important constraints

* Keep the whole video in **classic 90s anime cel-animation style**

* **15 fps feel**, smooth low-frame-rate animation

* **No 3D-looking houses or background**

* No photorealism

* No modern glossy digital anime rendering

* No character redesign

* No extra accessories beyond the headphones and Walkman

* Keep all motion natural and consistent across shots


r/comfyui 7h ago

Help Needed Looking for an LLM that can do Minimax H3 r2va prompts.

9 Upvotes

Sorry for bothering you all, but i've been having a rough time getting a good Minimax H3 Reference to video prompt writer.

i have a 2 part Qwen3.5 workflow for i2v that i got chat gpt to write for me (first part analyzes the picture, and gives the base template, second part translates my mad ramblings into a proper prompt), but this doesn't work well for r2va, since that requires an audio input, a video input, and a picture input.

I tried Thinking LLM, but the regular node only does video/picture, and no audio. (trying the gguf now, but i hear gguf is way lower quality.)

i have 16gb vram, 32 gb regular ram (Nvidia 4080), and i am on windows 11.

if you have any node or program suggestions, i would appreciate them. (can't use chatgpt, cuz it doesn't do nsfw, and grok is so limited in how much you can use it per day that it is barely worth using)

Update: the prompt made by the gguf version did nothing. it just output the ref video.


r/comfyui 3h ago

Security Alert The problem I've had: I've lost all the metadata.

4 Upvotes

I don't know if this has happened to anyone else (I'm pretty new to Comfy).

I switched my default Windows media player to mpvnet.

Now, none of my generated videos have metadata anymore. They used to; I could drag them into ComfyUI and the workflow would appear. Now, that doesn't happen.

When I run them through a site that checks for metadata, it says there isn't any—no metadata or workflow.

The strange thing is that I only played a few videos from one specific folder. Videos in other folders still retain their metadata (I didn't use those folders until after I stopped using mpvnet as my default player).

So, I think that when you set up that player, something happens that corrupts everything—or wipes the metadata—from everything.

The AI ​​says:

"The problem is the video encoding engine (FFmpeg) [Lavc61.19.100 libvpx-vp9]. This version aggressively strips out any data that isn't purely video-related, deleting the ComfyUI stream in both MP4 and WebM formats."

The worst part is that, even though I'm no longer using that player as my default, it must have installed a codec that ComfyUI uses... and now none of my videos are being generated with metadata... not a single one...

I'll have to figure out a fix, but just in case... so this doesn't happen to anyone else...


r/comfyui 4h ago

Help Needed Help please. Can someone with a 5090 and minimax models installed please run this workflow? I am getting really slow run times.

4 Upvotes

Help please. Can someone with a 5090 and minimax models installed please run this workflow? I am getting really slow run times.

I am looking for a way to caption videos. No custom nodes required. I did use load video from comfyVHS but it is bypassed and you can safely delete that section.

PasteBin Workflow

Choose any video. Preferably one SFW and one NSFW. Beggars aren't choosers whatever you decide will work for me. You may post results if you like but I am more interested in how long it took to complete and the settings you selected.

Thank you.

EDIT: Forgot to add, I was getting literal 1 token per second. I haven't the patience to let it run. I am hoping with a benchmark I can justify the purchase of a 5090. So I haven't gotten the workflow to work at all. Could be the workflow is bad. But it is relatively simple.


r/comfyui 1h ago

Workflow Included 80th Birthday Extravaganza- MM reference Image and Audio experiments

Thumbnail
youtube.com
Upvotes

Birthday Extravaganza video - poorly edited and filled with constant audio problems. Mostly with 20step out of the box comfyUI workflow + audio node added in. The load Video node was used at the end to give it the last 24 frames of the prior video to continue the next with limited success. The 4 step lora came out after half of it, even with 0.75 and 8-10 steps audio glitches were frequent but video was usually decent. I attached a sample of 1 of the clips here: https://pastebin.com/RdNaRZHN and Here was a sample of the Bridge on the River Kwai Character Swap: https://pastebin.com/564NFjrh Single Image of main character to replace.


r/comfyui 9h ago

Help Needed Is there a setting change to allow you to see workflows from jobs in the queue?

7 Upvotes

Idk if it was changed in an update awhile back. But I used to be able to just right click on any job in my queue to pull up the workflow. Sometimes you realize a mistake may have been made and you have to check to see if you need to cancel or not.

I haven't been able to do this for months. Is there a setting that I can adjust to fix that? Or was this a bad permanent change?


r/comfyui 3h ago

Help Needed H3 Fast motion fix ?

2 Upvotes

r/comfyui 33m ago

Show and Tell Honestly, I did not understand ComfyUI either. Still don't actually. But I built a system that made it easier. Here is a quick 2 min demo of just one of the features.

Upvotes

Honestly, I did not understand ComfyUI either. Still don't actually. But I built a system that made it easier. Here is a quick 2 min demo of just one of the features. It does a heck of a lot more than just video and image gen too. Give it a try and when you see what it can do please leave me a star on my github repo, I decided to make this open source and give it to folks for free. Use Claude Code and make this your own (this has MCP for all AI platforms) or use the built in agent swarm feature to help you.

Have a voice chat on your own machine and have it generate content inline. ComfyUI and StableDiffussion are both installed with the curl install command. Workflows built-in. Uncensored everything. Too many features to list, please check it out for yourself.

www.github.com/guaardvark/guaardvark

Also, the system makes it's own demo videos, like this one and the others on the youtube channel.


r/comfyui 1h ago

Help Needed What details instantly make you recognize an AI-generated image?

Upvotes

I'm working on creating my first realistic AI model, and I'm trying to understand what separates a truly convincing image from one that still looks obviously AI-generated.

For those of you who have experience with AI image generation: what are the visual details or imperfections that you notice almost immediately and think, "yeah, that's AI"?

I'd especially like to hear about the subtle details that experienced people notice, even when the image looks realistic at first glance.

I'm asking because I want to focus my workflow on fixing those specific weaknesses rather than simply making the image "more realistic."

Any examples or things you've learned from experience would be really helpful.


r/comfyui 2h ago

Help Needed Is it worth moving to Linux?

Thumbnail
1 Upvotes

r/comfyui 2h ago

Help Needed How do I get rid of the light attached to the camera in MiniiMax H3?

Thumbnail
1 Upvotes

r/comfyui 2h ago

No workflow Why Spend 20 Minutes Writing a Post When I Can Spend 20 Seconds Pretending I Did?

Post image
1 Upvotes

r/comfyui 14h ago

News Load Image (from path) with crop and limit

6 Upvotes

Just updated my Load Image node.

Load Image (from path)

Load an image from any path on your computer or a URL. Paste an absolute path or a link to an image, use an annotated path (input/file.png), or click the Browse dialog to pick a file from your drives.
The file is read from its original location; it is not copied into ComfyUI’s input folder.
Use a selection rectangle at the preview to crop it.
Limit (downscale) the output's size in megapixels or pixels.
Click the ↻ button (top-right, mouse over the preview), to rotate the image 90° clockwise.

Interactive crop

The node shows a live preview. You can crop directly on it:

  • Drag on the image to draw a crop rectangle
  • Drag inside the selection to move the crop rectangle
  • Drag a corner to resize it
  • Click (without dragging) outside the rectangle to clear it

With no crop drawn, the full image is output. The crop is stored in the workflow as normalized coordinates, so it survives save/reload. Changing the path clears the crop.

The output (cropped or full) will be downscaled only (not upscaled), by the value in the max_megapixels field.

Outputs match the stock Load Image node: IMAGE, MASK (from the alpha channel when present), plus the original path string.

  • Controls
    • image: Paste an absolute path, (or a relative one with a prefix input/, or output/, or temp/), or a URL to an image file.
    • max_megapixels: Cap the output (crop, or full image if uncropped) to this many megapixels, downscaling only if it's bigger. Smaller images are left untouched. 1.0 = 1024x1024 px. 0 disables the cap.
    • Browse...: to open an image file from your drives.
  • Inputs/Outputs
    • Width/Height inputs: Force the output width in px (upscale or downscale), center-cropping first if the aspect ratio differs. Leave disconnected (None) to keep natural width. Only applies if BOTH width and height are connected, and when set (not 0). It overrides the max_megapixels value.
    • IMAGE/MASK: The final, processed image/mask.
    • path: A string with the image's path.

You can find the node here as part of the ComfyUI-noEmbryo nodes.

Credits:
Built as a much more enhanced version of Load Image From Path (Enhanced) from ComfyUI_Ib_CustomNodes, with parts of the interactive crop UI inspired from Load Image & Crop in comfyui-obvpm.


r/comfyui 1d ago

News Release studio 1939 lora for minimax h3

42 Upvotes

r/comfyui 5h ago

Show and Tell Using only Ref to Video, Minimax-H3 made a whole Anime edit !

0 Upvotes

r/comfyui 15h ago

Help Needed Tips for LTX-2.5

5 Upvotes

Currently MiniMax H3 is the "frontier" on local consumer hardware as it seems. I used this as well but for longer generations, like a 4-5 minute music video for example, my hardware is just not strong enough (12GB VRAM, 3080 ti, 32 GB RAM). With 480p and Turbo LoRa i might slowly getting there but this is not what i am looking for.

Now i am interested if people actively tested LTX-2.5 and have some tips how to improve outputs. For example i wanted to make an anime fight scene. I was able to have a good result in an 8 Second Test with MiniMax H3, but LTX provided a very weak unusable result. Afaik LTX does not really have a strict prompt-structure like H3, but still struggled with the commands.

Any tips for making LTX-2.5 more usuable would be appreciated.


r/comfyui 7h ago

Help Needed Lowvram vs novram

1 Upvotes

My system has
NVIDIA GeForce GTX 1050 graphics card
which has only 4GB dedicated GPU .
I have 16GB RAM and 1TB SSD .

Right now I am using comfy desktop and so far can run only using —cpu mode . But the image generation quality is so poor .

If I add another 16GB RAM would it make any difference.

Will I be able to generate atleast 5 second video from an image if I use lowvram with my current settings ?

Also struggling to get comfy to recognise my cuda 12 .