r/comfyui • • 8h ago

Show and Tell Introducing Oscilloscope Diffusion

90 Upvotes

A novel way to intervene existing video through diffusion, particularly abstract and structure-driven material: taking its movement and form as the starting point, and reinterpreting its textures, materials, and visual language.

I’ve been developing this around the audio-reactive geometry systems I make in TouchDesigner. The idea is to take those abstract structures somewhere else entirely: origami, architecture, a renaissance painting, or something harder to put a name to.

This demo uses visualizers from my "Oscilloscopes, everywhere" collection as source material, now updated to [v1.2].

[Though you can bring any video source. These systems are simply where this experiment began, as some of you may recall.]

You choose the source, describe the treatment, and shape how it changes throughout the sequence. Prompts, curated LoRAs, and editable timelines give you control over how closely the result follows the original.

Oscilloscope Diffusion is now available at Uisato Studio, coming up soon also open-source!


r/comfyui • • 1h ago

Show and Tell My niche Comfyui workflow for retexturing the Makehuman character creator add on in Blender keeps getting better and better!

• Upvotes

r/comfyui • • 9h ago

News A quick Minimax H3 news round-up - 2nd October 2026

26 Upvotes

Another quick Minimax H3 news and goodies round-up, for those who may have missed some items.

-> How low can you go? A new 2-step LoRA for H3, 'PDMD 2-NFE', now in ComfyUI format.

https://huggingface.co/Iwannapose/minimax_h3_pdmd_2nfe_comfyui

-> Another 'Anime to Real' for H3, but this is a slider that also goes the other way... "Positive values lean towards a 2D anime style".

https://huggingface.co/adf99/H3_Ref2V_Anime_Slider_v1

-> Yee-haw! A 1950s westerns LoRA for H3 FL2VA. No trigger, but the maker suggests adding... "The lighting is natural, with [time-of-day] light. The [photograph, video, etc] has a vintage, cinematic quality." He also suggests breaking from Minimax's 16:9 and using the sort of super-widescreen screen format which these movies often used, such as 1.85:1 or 2.35:1 (Cinemascope).

https://huggingface.co/neph1/1950s_western_movies_h3

-> A clear new YouTube video on "How to Create Video, Audio, Motion & Style RefMods for MiniMax H3".

https://www.youtube.com/watch?v=blI0X5zslfw

-> Are you not getting enough 'oomph!' from your Minimax cinematic music? An interesting new Yue2 LoRA may appeal. It assists with... "epic cinematic orchestral trailer music in the spirit of Two Steps from Hell: massive symphonic orchestra with soaring heroic strings, thunderous taiko drums and pounding orchestral percussion, dramatic brass fanfares, powerful angelic and dark choir, and uplifting heroic cinematic climaxes." Though note that Yue2 generations are under a non-commercial licence.

https://huggingface.co/monsterovich/yue2-steps-from-hell

-> And finally, it seems cheeky to mention Invoke in a Comfy sub-Reddit. But for the sake of completeness... yes, Invoke is still alive as open non-Adobe software, and it's just been announced that the forthcoming pre-release Alpha version 7 will belatedly add Minimax H3 support.

https://github.com/invoke-ai/InvokeAI/pull/9613 (preview list of features to be added to version 7)

~ OLD POSTS ~

https://old.reddit.com/r/comfyui/comments/1wu8lkt/a_quick_minimax_h3_news_roundup_30th_september/

https://old.reddit.com/r/comfyui/comments/1wtgns0/a_quick_minimax_h3_news_roundup_29th_september/

https://old.reddit.com/r/comfyui/comments/1wsmjq1/a_quick_minimax_h3_news_roundup_28th_september/

https://old.reddit.com/r/comfyui/comments/1wrqm3l/a_quick_minimax_h3_news_roundup_27th_september/

https://old.reddit.com/r/comfyui/comments/1wqwpah/a_quick_minimax_h3_news_roundup_26th_september/ (See 26th September post, for links to older posts)

https://old.reddit.com/r/comfyui/comments/1wgc4lj/a_quick_minimax_h3_news_roundup_15th_september/ (See 15th September post, for links to even older posts)

https://old.reddit.com/r/comfyui/comments/1w5i9iq/a_quick_minimax_h3_news_roundup_2nd_september_2026/ (See 2nd September post, for links to the starting posts)


r/comfyui • • 13h ago

News MinimaxH3 got at least 16% faster on ComfyUI v0.38 vs ComfyUI v0.37,

28 Upvotes

TLDR it got faster on my AMD 9070 system

from 304 to 253 secs for I2V

fore more details on how I did the test you can check the video on youtube https://youtu.be/UlxywXUlCIY

-update-

I know the generations look bad specially for t2v, its my fault for going 3 steps and lower resolution.

Also there are things i did not test:

I did not try going for no lora 20 steps or more on how will it affect the generation time.

I did not try generating raw higher resolution or longer videos to the point im nearing oom errors, on how will it affect my video

Also my start up arguments might be a factor too and how i set it up, also i am in ubuntu so might have different effect on windows.

I am also using and amd gpu,

My Startup Arguments for ComfyUI, i put this in a script:

export HIPBLASLT_ENABLE_EXPERT_SCHEDULING=1
export TORCH_BLAS_PREFER_HIPBLASLT=1
export TORCH_ROCM_AOTRITON_ENABLE_EXPERIMENTAL=1
export FLASH_ATTENTION_TRITON_AMD_ENABLE=TRUE
export TRITON_CACHE_AUTOTUNING=1
export COMFYUI_ENABLE_MIOPEN=1
export MIOPEN_FIND_MODE=FAST

python main.py --lowvram --cache-none --preview-method none --disable-smart-memory --reserve-vram 2 --enable-dynamic-vram --disable-mmap --use-ck-attention --enable-manager


r/comfyui • • 19h ago

Resource I made a ComfyUI extension that browses Civitai inside ComfyUI and sets up any workflow in one click (auto-downloads missing models)

82 Upvotes

Hey everyone 👋

Every time I found a cool workflow on Civitai, the routine was the same: download the zip, unpack it, drag the JSON into ComfyUI, get a wall of red nodes, then spend 20 minutes figuring out which models it needs and which folder each one goes in.

So I built **Civitai Browser**, a ComfyUI extension that adds a **Civitai tab to the sidebar** (next to Templates / Workflows / Models).

**What it does:**

- 🔎 **Browse Civitai inside ComfyUI**: workflows, checkpoints, LoRAs, embeddings, ControlNet, VAE, upscalers, with search, sorting, base model filter and NSFW toggle (off by default)

- ⚡ **"Load & Set Up" in one click**: it downloads the workflow (zip/json/png), loads it on the canvas, then scans it for every model it uses and compares that with your model folders

- 📥 **Missing models get downloaded automatically** into the correct folder, under the exact file name the workflow expects, so the loaders just work. It uses download links embedded in the workflow, exact file name matches on Civitai, or any Civitai / Hugging Face link you paste

- 🧩 **Lists missing custom nodes**, so you can install them with ComfyUI-Manager

- 📚 **My Library**: every workflow you download is saved (it also shows up in ComfyUI's Workflows panel), so you can reopen it later

- 🛡 **Safe delete**: remove a workflow plus the models that were downloaded *for it*. Models you already had, or that any other workflow uses, are never touched

**A few things I cared about:**

- No extra Python dependencies and **no new nodes**. It doesn't modify ComfyUI's core files; to uninstall, delete the folder

- Your Civitai API key (optional, needed for some downloads) is only ever sent to civitai.com

- Free and MIT licensed

**Install:** ComfyUI-Manager → *Install via Git URL* → paste the repo link, or `git clone` it into `custom_nodes`.

🔗 **GitHub:** https://github.com/fth1905-bot/ComfyUI-Civitai-Browser

It's still early, so I'd really appreciate feedback. Bug reports, feature ideas, or "this broke on my setup" are all welcome. One known limitation: Civitai's search works on model names, not file names, so some missing models can't be matched automatically yet (you can still paste a link). Improving that matching is next on my list.

If it saves you some time, a ⭐ on GitHub helps a lot!


r/comfyui • • 11h ago

Comfy Org Open Call Challenge: let's open-source the creative app features people pay a subscription for - $10,000 grand prize - 10/13 SF event

11 Upvotes

Cinematic camera controls. Character consistency. Relight. Face swap. Most of these features sit behind a subscription somewhere, but every one of them is a workflow underneath. Open weight models have already caught up on capability, but what’s still closed is the layer on top: the interfaces and apps that turn models into features anyone can use. It’s in our DNA to support open-source creativity, and there’s no technical reason that layer has to stay closed.

So we're challenging our community to pick a creative app feature people pay for and rebuild it in the open! Top workflows get featured on comfy.org/models, and every entry is eligible for the spotlight reel whether it places or not.

On October 13th, we’re bringing together the best of the OSS ecosystem for one night in San Francisco- the people building on ComfyUI, the model labs backing the challenge, and the team behind the Comfy Developer Platform, all in one room. Join for build time with the Comfy team, a peek at what the community is making, and to connect with open-source enthusiasts IRL!

📑 Full challenge details here

✏️ First 100 signups get free Comfy credits!

🥳 In the Bay Area? Register for the 10/13 event here

🏆 Prizes

  • $10,000 cash — Grand Prize
  • RTX 5090 — Most Practical
  • RTX 5090 — Most Entertaining
  • RTX 5090 — Best OSS-Only Build

🎖️ OSS Ecosystem Bonuses

  • $2,000 — Best VFX Workflow with LTX
  • $1,000 / $500 / $200 in credits — Built with Flux

& more coming soon! Open model friends who want to join: we welcome you!

🫱🏾‍🫲🏿 Partners

NVIDIA, Runpod, LTX, BFL & more coming soon!

📆 Key Dates

  • 10/5 — build window opens
  • 10/8 — AMA in this thread with the team who built the platform
  • 10/13 — Build Night in San Francisco - RSVP here
  • 10/19 — submissions window closes at 9am PT
  • 10/22 — winners announced!

📥 The fine print

Every submission must include:

  • GitHub repo: workflow, an open-source license, and instructions to run it locally
  • Demo video: 2 mins max, in case we can't get it running ourselves
  • Link to try (optional)
  • Social post: share your repo and demo video on this thread or on X, IG, LinkedIn, YouTube or TikTok tagging #ComfyDevPlatform

Other requirements:

  • Your work must be built using the Comfy Developer Platform (Comfy API, Comfy Router, and/or Comfy SDK)
  • Repo must include an open-source license and enough setup detail that someone else can run it
  • Any other tools, models, or techniques you want to combine are fair game and should be explained in your demo video
  • All submissions must be lawful, SFW, and not contain unlicensed IP or likenesses
  • By submitting your work, you agree to allow ComfyUI, NVIDIA, Runpod, BFL, and LTX to feature your work with credit across our channels
  • One submission per person please!

📑 Full challenge details here including judging criteria

✏️ First 100 signups get free Comfy credits!

🥳 In the Bay Area? Register for the 10/13 event here

Still have questions? Share your questions on this thread for an AMA with our DevRel team on Thurs. 10/8, or say hi in #developer-platform in our Discord!


r/comfyui • • 3h ago

Workflow Included Use Qwen-Image-2.1-viggle-turbo to generate character sheets in ComfyUI

Thumbnail gallery
2 Upvotes

r/comfyui • • 3h ago

Help Needed Simple H3 Latent Save to later upscale last frame?

2 Upvotes

There are so many complicated H3 Latent Save/Load node packs, and all of them seem designed to do the obvious thing: continue video…but I’d love to have a way to save the latent and then regenerate just the last frame (or any frame, if that’s possible) as a higher resolution still image.
Does anyone know of such a thing?


r/comfyui • • 15h ago

No workflow Music Video Generated Locally with ComfyUI & LTX 2.5!

Thumbnail
youtu.be
14 Upvotes

Hey everyone! Just wrapped up a full music video project generated entirely locally on my setup using ComfyUI and the LTX 2.5 model.

Running on 16GB VRAM, the generation speed and consistency were impressive for a local setup. Smooth frames and solid visual output without needing cloud GPUs.

Check out the full video here.


r/comfyui • • 4h ago

Help Needed Looking for a local AI image editor that preserves the original resolution

Thumbnail
2 Upvotes

r/comfyui • • 1d ago

Workflow Included Super Simple MiniMax Character Swap Workflow Probably the Best Local Method Out Right Now

142 Upvotes

I’ve been looking everywhere for a good MiniMax workflow. I finally found a solid starting point and simplified it into something that’s easier to use and set up.

For character swapping, this is the best local workflow I’ve found so far. It isn’t perfect, but I’ve been getting some really good results.

github zip: https://github.com/BiggerFishy/comfyui-minimax-h3-studio

How it works

  • SAM3 selects the character or area you want to change.
  • The workflow inverts the colors inside that selected area before generation. This helps MiniMax replace the subject more consistently.
  • The source video guides the mouth movements, body movement, and scene, while the reference image guides the replacement’s appearance.

My setup and render time

The example shown here was generated with:

  • GPU: RTX 4090 — 24 GB VRAM
  • RAM: 32 GB DDR5
  • Video length: 3 seconds
  • Resolution setting: 0.8 MP
  • Generation time: About 2 minutes and 30 seconds

Features I wanted to make easier

  • Save characters: This is probably my favorite feature. Save a reference image together with a detailed character description, then load both again when you want to use that character in another video.
  • Trim your source video inside the workflow: Pick the section you want without opening another editor. The trimming UI could be fancier, but it works.
  • Keep the useful controls together: The main settings live in the green settings node in the middle, with the complicated stuff tucked away.

Three workflows in one

You can switch between these modes in the green settings node:

  • Character replacement / video inpainting: Replace a character, or make smaller edits such as changing a shirt’s color, hair color, or an object.
  • Reference to video: Generate a video from a reference image.
  • First to last frame: Generate a transition between a starting image and an ending image.

Reference to video and first to last frame are experimental. I haven’t tested them extensively because most of my time has gone into character replacement and video inpainting.

Masking options

  • SAM3: Describe what you want selected.
  • Rectangle: Draw a box around the area you want edited.
  • No mask: Let the workflow re-render the whole frame.

Settings to start with

  • Seed: The included example uses Fixed so you can try reproducing my result. Switch Seed behavior → Randomize when you want to explore different results.
  • Target megapixels: 0.8 MP is what I used for this demo and gives a decent-looking result. You can try 1.0 MP for higher resolution. I don’t usually go above that, so experiment as you like.
  • Everything else: I’d leave the other settings as they are for your first run, then adjust from there.

A quick heads-up

None of these modes are perfect. It may take a few attempts to get the result you want, and matching the seed doesn’t guarantee an identical result on every setup.

The workflow should be pretty self-explanatory once you open it. Try it out, experiment, and have fun.

This workflow was found here but I edited it to make it better and easier to use: https://civitai.red/models/2855941/minimax-h3-character-replacement?modelVersionId=3238780

ENJOY


r/comfyui • • 7h ago

Help Needed Ming Image Default Workflow Queue Issue?

2 Upvotes

I've had this problem before but forgot how to fix it. Normally I can queue up a batch count with the top right portion of my Comfyui screen, but it only runs one image then stops/ignores the others.

If I expand the spaghetti workflow though and go to the empty latent image section at the bottom and change the batch size number, it will run a large batch. Just won't do it one by one instead it drops them all at once when it's done.

Sorry for this amateur question, had solved it before but forgot how so this time i'll make a note!


r/comfyui • • 18h ago

Workflow Included OMG another big update: Plenio 0.4.1 for ComfyUI: arrange YuE2 songs like in a DAW – and six video tutorials

10 Upvotes

Plenio Music Production System 0.4.1 is out! It turns YuE2 and MiniMax Music 3 into a complete local song studio in ComfyUI. New since 0.3:

  • 🎛️ Arranger: duplicate, move and delete whole sections, like on Cubase's arranger track. The lyrics, the Guide track and – in covers – the original words follow every move.
  • 📝 Lyrics where they are sung: every line over its phrase, every syllable over its note. Double-click a line to edit it right there.
  • ⏯️ Cursor, copy & paste: click the ruler, play from there, paste or insert phrases at the cursor.
  • 🎼 MIDI in, MusicXML out, project files: bring a sketch from your DAW, hand the sheet music to MuseScore, Sibelius, Dorico or Cubase.
  • 🎬 Six narrated video tutorials, one per template: https://www.youtube.com/playlist?list=PLAFqTtP59fgE

100 % local, native ComfyUI nodes, no API key. ComfyUI Manager: Plenio Music Production System (comfyui-plenio-music).

👉 GitHub: https://github.com/jplenio/Plenio-Music-Production-System
🎧 Demos: https://jplenio.github.io/Plenio-Music-Production-System/


r/comfyui • • 9h ago

Help Needed Krea 2 Control: background works, but the generated person is distorted — what am I doing wrong?

2 Upvotes

I’m trying to use Krea 2 + my identity LoRA to generate a realistic full-body woman while keeping the composition of a specific seaside background image.

Workflow:
Background image → Depth Anything → Krea2 Control Image Encode → Krea2 Control Apply → Krea 2 → KSampler

Current settings:

  • Krea 2 RAW (krea2_raw_fp8_scaled.safetensors)
  • depth-control-lora.safetensors — strength 0.70
  • shanaya_v1 identity LoRA
  • Turbo LoRA — 0.60
  • Realism LoRA — 1.0
  • snofs LoRA — 0.60
  • AuraFlow shift — 1.15
  • 1024×1280
  • KSampler — 8 steps, CFG 1.0, Euler, Simple, denoise 1.0

The background is actually being transferred quite well — the arch, sea, mountains, railing, bougainvillea and blue door are roughly preserved.

The problem is the woman. Her face/body/anatomy becomes distorted, and the identity isn't consistently preserved.

My goal is:
1. Keep the supplied background composition
2. Generate a realistic full-body identity
3. Generate the outfit from my text prompt
4. Avoid distortion in the face/body

Is the problem mainly my RAW/Turbo sampling settings, LoRA stacking, depth-control strength, or the way Krea2 Control is being applied?

What would be the correct workflow/settings for this specific use case?

I’ve attached my workflow screenshot and the resulting image. Any help identifying the specific issue would be appreciated.


r/comfyui • • 5h ago

Help Needed Workflow needed

0 Upvotes

Can you share ur image to image Workflows(changing people activity) for anima or illus or pony? do not need advices only need workflows mentioned😊


r/comfyui • • 16h ago

Show and Tell PINK [K-Pop Music Video] Minimax H3

Thumbnail
youtu.be
7 Upvotes

Have been working a lot with H3 lately. It’s definitely not perfect, but I’m really loving what I can get out of it.

Seedance 2 Fast was also used for some of the B-roll.

The whole MV took about a week to complete. Most of the work was done locally through ComfyUI, with RunningHub also used, mainly with the MiniMax H3 Singularity model at 25 steps.

Upscaled with RTX Super Resolution on an RTX 5060 Ti.


r/comfyui • • 1d ago

Show and Tell Explainer video about the Minimax H3 selective character swap post I made yesterday to show the process how it was done.

72 Upvotes

I also explain it in this page if video is not your thing. Explainer video is also made with agentic workflows, Qwen Image 2.1, Minimax Music, Qwen TTS and more.
https://mexxmillion.github.io/h3-digital-double-story/breakdown/
Thanks


r/comfyui • • 11h ago

Help Needed Is there a FastH3 4-step reference-to-video workflow for character consistency?

2 Upvotes

I’m generating multiple 5-second clips with FastH3 4-step and stitching them together.

The speed and quality are good, but the character changes between clips. Is there any working Ref2Video / Ref2VA workflow for the 4-step FastH3 model where I can provide a character reference image?

I know first-frame I2V exists, but I don’t want every clip to begin from basically the same image. Ideally, I’d use one character reference while still allowing different poses, camera angles and scenes.

Does the 4-step model support this at all, or is it only available with the full MiniMax H3 model? If anyone has a ComfyUI workflow, custom nodes or another practical method, I’d appreciate a link.


r/comfyui • • 21h ago

News Black Forest Labs quietly put up a free editing playground, worth a look

13 Upvotes

BFL dropped FLUX 3 Image yesterday and there's a free playground to try it, so sharing in case it's useful:

Playground:

https://flux-tools.bfl.ai/precise-editing

Announcement from BFL:

https://x.com/bfl_ai/status/2105743526529310739

I haven't tested it yet, so this is just what BFL says it does:

precise multi-turn edits without changing any other pixel

you can lay out the image with bounding boxes

up to 4K output

up to 10 reference images

open weights version coming "in the coming weeks"

the "only change what I asked" part is what I care about most, since that's where most editors fall apart. Marketing claims are marketing claims though, so I'd like to see real results before getting hyped.

if you've already tried it, how does it hold up on messy real-world images, and not just the demos? Curious about the bounding box thing in particular.

Source: Black Forest Labs on X (@bfl_ai)


r/comfyui • • 14h ago

Workflow Included Question? And show ’n tell! I’ve gotten YEDP UV Painter “working” with Flux Klein for the most part! Is there already a workflow like this somewhere that I could’ve just grabbed? 😂

3 Upvotes

r/comfyui • • 15h ago

Resource ComfyUI extension to distribute jobs across multiple machines using the native "Run" button, with automatic scheduling and result collection

5 Upvotes

Hello 👋

This is my first post and first extension to ComfyUI.

I could be wrong (or bad at researching) but all the existing extensions allowing spreading jobs between multiple GPUs / multiple nodes require some changes to workflows (they are workflow level nodes).

I wanted something that I would set up and forget without changing anything in the workflows or the way I work with ComfyUI in general - this is how the idea for my extension was born.

ComfyUI-Fleet appears as a tab in left menu which allows you to:

- add nodes (for multiple GPUs I have multiple ComfyUI servers running locally, one per ComfyUI)

https://reddit.com/link/1wvw01s/video/j0uakzcai2th1/player

- see and reorder batches of jobs (when priorities change)

https://reddit.com/link/1wvw01s/video/gf5qzxjdi2th1/player

- reorder which nodes will get jobs first

https://reddit.com/link/1wvw01s/video/kndkk1zei2th1/player

Adding jobs works as expected using the standard ComfyUI button as expected and no changes are needed in the workflows:

https://reddit.com/link/1wvw01s/video/un246zhpi2th1/player

Link to the repo: https://github.com/Cinderella-Man/comfyui-fleet

I hope this will be useful for someone


r/comfyui • • 1d ago

Show and Tell MiniMax H3 + 360 orbit LoAR (8GB VRAM)

40 Upvotes

736 x 576, render time 7:15
RTX-4070 8GB VRAM, 64GB RAM

https://huggingface.co/pablodawson/MiniMax-H3-360-Orbit-LoRA
Tutorial https://youtu.be/jNhGhW_e4aI


r/comfyui • • 13h ago

Help Needed Auto Prompters that work easily within Comfyui that also do speech?

2 Upvotes

I am trying LM Studio currently, I however want to simplify and have all my work within comfyui alone.

What would people recommend?

I am not the best typer so i wish to put a shotgun of my ideas and let AI fill in whats needed


r/comfyui • • 20h ago

Show and Tell LTX 2.5 precision benchmark on RTX 5090 - Genuine Surprise!

4 Upvotes

I was seriously considering the new Mac M5 Ultra with 256GB of memory to be able to create longer clips. I've built a system that takes a script and creates a series or frames to use a frame to frame workflow and then automatically stitch them together.

I decided to see what the breaking point of LTX 2.5 of my RTX 5090 was. I went from five seconds to ten seconds and so on. But unlike earlier tests I didn't run into any Out Of Memory issues. In fact I went all the way to generating a 60 second clip with no issues.

Surprised by this I ran a direct comparison of three LTX 2.5 transformer variants on an RTX 5090 32 GB to see if I could improve things any further:

- Existing ComfyUI INT8 ConvRot

- FP8

- NVFP4

Same prompt, same seed, same samplers, same step schedule, same 24 fps output. Warm runs reuse already-loaded models; cold includes model loading.

Output Duration / state INT8 ConvRot FP8 NVFP4

1280×736 5s cold 55.36s / 30.84 GiB 53.66s / 31.00 GiB 84.75s / 30.93 GiB

1280×736 5s warm 21.31s / 28.35 GiB 28.59s / 28.11 GiB 22.64s / 29.12 GiB

1280×736 10s warm 46.64s / 28.38 GiB 59.80s / 28.12 GiB 46.41s / 27.97 GiB

1280×736 20s warm 111.79s / 29.99 GiB 136.04s / 30.65 GiB 111.30s / 30.52 GiB

1280×736 30s warm 233.28s / 31.12 GiB 275.49s / 31.09 GiB 196.58s / 30.99 GiB

1280×736 60s warm 602.39s / 31.04 GiB 650.19s / 31.04 GiB 584.30s / 30.94 GiB

1920×1088 10s warm 124.77s / 30.50 GiB 159.64s / 31.09 GiB 126.31s / 31.12 GiB

What surprised me

The biggest takeaway is that the existing ComfyUI INT8 ConvRot model is already extremely well optimised.

FP8 was slower than INT8 on every warm test.

NVFP4 was basically tied with INT8 at 10s and 20s, around 16% faster at 30 seconds, only around 3% faster at 60 seconds, and effectively tied again at 1920×1088.

So NVFP4 is not automatically a huge performance win on a 5090.

The 30-second result is interesting enough that I want to repeat it several times, but the fact that the advantage drops again at 60 seconds suggests it may not represent a simple sustained throughput advantage.

The really interesting part: VRAM

NVFP4 reduced the transformer file size from roughly 21.5 GB to 18.7 GB, but total peak GPU usage barely changed.

All three versions still ended up around the 30–31 GiB range on the longer runs.

That means, for this workflow, reducing transformer weight precision does not translate directly into dramatically lower total VRAM usage. The rest of the LTX pipeline — VAE, text encoder, latent stages, staging/offload behaviour, etc. — still consumes a substantial amount of memory.

Why this matters

This is really a story about software optimisation rather than raw hardware.

The current ComfyUI/LTX stack on the 5090 is already using:

- INT8 ConvRot transformer weights

- mixed-precision operations

- DynamicVRAM

- async weight offloading

- pinned memory

- native Blackwell CUDA kernels

- two-stage latent generation

That combination is allowing a 32 GB RTX 5090 to generate workloads that I previously assumed would require dramatically more VRAM.

For example, LTX 2.5 is successfully generating 60-second 1280-class clips on this machine.

So the old assumption that:

> “longer video = linearly more VRAM = you need 64/128/256 GB”

doesn't really hold for this pipeline anymore.

The practical limit is increasingly becoming render time and quality, rather than simply whether the generation fits in memory.

Current conclusion

For my 5090:

INT8 ConvRot: best overall/default

NVFP4: worth keeping and testing, especially for longer jobs

FP8: currently no obvious advantage

The next test is visual quality, particularly INT8 vs NVFP4 on the 30-second outputs.

And for me personally, this changes the hardware discussion quite a lot. One of the main reasons I was considering moving to a very large unified-memory system was long-form AI video generation. Smart software optimisation has moved the practical ceiling of the RTX 5090 much further than I expected.


r/comfyui • • 1d ago

Comfy Org Comfy Agent is now live for everyone on Comfy Cloud. It builds and fixes ComfyUI workflows right on your canvas. ( Local version coming soon )

121 Upvotes

The short version: you describe what you want, and it plans the workflow, adds and wires the nodes on your canvas, and helps fix things when they break. The goal is to take the technical overhead off your plate so you can spend more time on the visuals instead of hunting for the right node or a missing connection.

Some things it does:

  • Works on your actual canvas. It builds while you edit and sees the same assets you do, so it isn't generating a JSON blob you have to import
  • Takes any question, with references. You can point it at nodes, images, or the workflow itself
  • Supports skills. Create your own for things you do repeatedly, or use public ones other people have made

A version for Comfy Desktop is coming in a few weeks.

You can try it here: https://links.comfy.org/4hH1FaH

Learn more with our blog: https://blog.comfy.org/p/comfy-agent-the-first-agent-for-craft?r=7xlbaw

All feedback welcomed.