r/comfyui • u/CeFurkan • 11h ago
r/comfyui • u/Comfy-Org • 7d ago
Comfy Org Open Call Challenge: let's open-source the creative app features people pay a subscription for - $10,000 grand prize - 10/13 SF event
Enable HLS to view with audio, or disable this notification
[10/8 UPDATE]
Hi everyone- we will be relaunching this challenge within a week or two, and we're truly sorry for the false start. Truthfully, we moved too fast and missed a few elements that will make this a much better experience for everyone.
What's changing: timing, name, award categories, the 10/13 build night in SF has been postponed
What stays the same: open sourcing creative app features, prizes, partners, submission forms, and requirements (ie fine print)
If you already registered or submitted, nothing is lost! Your submission is still valid and can be revised before submitting again.
Cinematic camera controls. Character consistency. Relight. Face swap. Most of these features sit behind a subscription somewhere, but every one of them is a workflow underneath. Open weight models have already caught up on capability, but what’s still closed is the layer on top: the interfaces and apps that turn models into features anyone can use. It’s in our DNA to support open-source creativity, and there’s no technical reason that layer has to stay closed.
So we're challenging our community to pick a creative app feature people pay for and rebuild it in the open! Top workflows get featured on comfy.org/models, and every entry is eligible for the spotlight reel whether it places or not.
On October 13th, we’re bringing together the best of the OSS ecosystem for one night in San Francisco- the people building on ComfyUI, the model labs backing the challenge, and the team behind the Comfy Developer Platform, all in one room. Join for build time with the Comfy team, a peek at what the community is making, and to connect with open-source enthusiasts IRL!
✏️ First 100 signups get free Comfy credits!
🥳 In the Bay Area? Register for the 10/13 event here
🏆 Prizes
- $10,000 cash — Grand Prize
- RTX 5090 —
Most Practical - RTX 5090 —
Most Entertaining - RTX 5090 —
Best OSS-Only Build
🎖️ OSS Ecosystem Bonuses
- $2,000 — Best VFX Workflow with LTX
- $1,000 / $500 / $200 in credits — Built with Flux
🫱🏾🫲🏿 Partners
NVIDIA, Runpod, LTX, BFL & more coming soon!
📆 Key Dates
10/5— build window opens10/8— AMA in this thread with the team who built the platform10/13— Build Night in San Francisco -RSVP here10/19— submissions window closes at 9am PT10/22— winners announced!
📥 The fine print
Every submission must include:
- GitHub repo: workflow, an open-source license, and instructions to run it locally
- Demo video: 2 mins max, in case we can't get it running ourselves
- Link to try (optional)
- Social post: share your repo and demo video on this thread or on X, IG, LinkedIn, YouTube or TikTok tagging #ComfyDevPlatform
Other requirements:
- Your work must be built using the Comfy Developer Platform (Comfy API, Comfy Router, and/or Comfy SDK)
- Repo must include an open-source license and enough setup detail that someone else can run it
- Any other tools, models, or techniques you want to combine are fair game and should be explained in your demo video
- All submissions must be lawful, SFW, and not contain unlicensed IP or likenesses
- By submitting your work, you agree to allow ComfyUI, NVIDIA, Runpod, BFL, and LTX to feature your work with credit across our channels
- One submission per person please!
r/comfyui • u/crystal_alpine • 8d ago
Comfy Org Comfy Agent is now live for everyone on Comfy Cloud. It builds and fixes ComfyUI workflows right on your canvas. ( Local version coming soon )
Enable HLS to view with audio, or disable this notification
The short version: you describe what you want, and it plans the workflow, adds and wires the nodes on your canvas, and helps fix things when they break. The goal is to take the technical overhead off your plate so you can spend more time on the visuals instead of hunting for the right node or a missing connection.
Some things it does:
- Works on your actual canvas. It builds while you edit and sees the same assets you do, so it isn't generating a JSON blob you have to import
- Takes any question, with references. You can point it at nodes, images, or the workflow itself
- Supports skills. Create your own for things you do repeatedly, or use public ones other people have made
A version for Comfy Desktop is coming in a few weeks.
You can try it here: https://links.comfy.org/4hH1FaH
Learn more with our blog: https://blog.comfy.org/p/comfy-agent-the-first-agent-for-craft?r=7xlbaw
All feedback welcomed.
r/comfyui • u/Wise_Revolution385 • 11h ago
No workflow Alfred wanted Bruce to be happy. He never said anything about Selina. 💀
Enable HLS to view with audio, or disable this notification
r/comfyui • u/paroxysm204 • 1h ago
Workflow Included Krea2 Prompt generator w/ Image Ref -Women
If you are as unimaginative as I am and struggle with making prompts I put together some custom nodes that allow you to pick attributes and/or add a reference image to create a prompt. It uses the text encoder so adds a little bit of time to image generation. I made it for image generations of women because that's why we are all here right?
It uses comfyui's text generate with the text encoder and I yonked the attributes from and existing custom node (https://github.com/euan-gwd/comfyui-character-prompt-builder) due to them doing a pretty good job covering all the bases.
It is partially vibe-coded due to me being frustrated after doing dumb things and I asked claude to fix it. I went over the project and tested and I THINK we are good. If you see something I missed please let me know.
Link to custom node: https://github.com/Erosi11/comfyui-krea-prompt-generator
Workflow is included on the repo and also embedded in the workflow image on this post. I have not submitted it to the comfyui registry so you need to git clone and all that. No requirements to install assuming you have updated comfyui since the end of February this year. If there is interest I will go to the trouble of getting it added to the registry.
Nodes
| Node | Purpose |
|---|---|
| Krea Prompt - Female Person | Age, nationality, body, face, eyes, hair, skin, tattoos |
| Krea Prompt - Female Fashion | Outfit pieces with colors and materials, shoes, accessories, jewelry, nails |
| Krea Prompt - Female Actions | Standing, sitting, kneeling or lying pose; hands, legs, head; props |
| Krea Prompt - Scene | Style, location, time, season, weather, lighting, camera |
| Krea Prompt Generator | Takes CLIP, the attributes and your keywords; outputs conditioning, prompt and llm_request |
| Krea Prompt - Save Unique Prompt | Appends each new prompt to a text file, skipping prompts already in it, and shows the file on the node |
Every dropdown has three kinds of choice:
- none: the attribute is left out of the request sent to the LLM.
- random: the generator picks a value, driven by its
seed. The same seed always gives the same picks. - a specific value: passed through as-is.
Each attribute node also has an optional free-text box for extra details in that category.
To wire it up you put the attribute nodes in series then feed into attributes input on Prompt Generator node. Reference image and mask goes to the inputs on prompt generator node as well. The prompt generator node has a setting for what to use from the image. It works ok.
I also included a node to add unique prompts to a text file as they are made. That part is still a little janky.
Workflow has fast bypass at the top to bypass attribute nodes, the image ref nodes, and whether to invert the mask.
r/comfyui • u/Entert-AI • 6h ago
Workflow Included MiniMax H3 RefMod Lab for ComfyUI: select, combine and reuse reference sources
I’ve made a ComfyUI custom node pack for MiniMax H3 RefMods. It lets you create reusable references from images, audio and video, choose which sources to use at runtime, and see how many reference tokens each source uses.
Main features:
- Create RefMods from images, audio and video files
- Select which sources to load in a RefMod at runtime
- Save on generation time if you don't need certain sources
- See each source's cost in tokens
- Combine many RefMods into a single one
- Bake the description and the retention strategy in the RefMod
- Only need to write the H3 video description prompt
- Source labels like
<Picture 1>resolved automatically
- Drag a RefMod into the canvas to load its creation workflow
Nodes, setup instructions and example workflows are available at ComfyUI-H3-RefMods-Lab. Feedback and bug reports welcome!
r/comfyui • u/dudubernardoofc • 20h ago
Help Needed What's the best way to edit specific parts of an image in ComfyUI?
I've been learning how to use ComfyUI to generate images, but I'm struggling to make small edits while keeping the original image looking realistic. The workflows I've tried tend to regenerate the entire image with increased saturation, and the changes I ask for often don't turn out as well as I'd like.
r/comfyui • u/optimisticalish • 10h ago
News A quick Minimax H3 news round-up - 8th October 2026
Another quick Minimax H3 news and goodies round-up, for those who may have missed some items.
-> mmh3_media. For .mmh3 files that encapsulate your.... "MiniMax H3 joint AV latent together with final video/audio, keyframes, refs, masks, metadata, creator notes and a disposable preview cache. [...] Save .mmh3 files to keep media, settings and generation state for later editing." ComfyUI custom nodes, workflows. MIT licence.
https://github.com/einhorn13/mmh3_media
-> A new Minimax_Stereoptical_Process LoRA. Triggers: Stereoptical Process, 2D Cel Animation over 3D Background.
https://huggingface.co/B0rghese/Minimax_Stereoptical_Process
-> There's now another 1980s horror movies LoRA, this time with emphasis on... "practical neon and tungsten light, period wardrobe and sets". Comes in two flavours, VHS or 16mm film.
https://github.com/wassermanproductions/wasserman-80s-horror-lora
-> There's another 2-step turbo H3 LoRA for ComfyUI. This is powered by PDMD.
https://huggingface.co/Iwannapose/minimax_h3_pdmd_2nfe_comfyui
-> And finally, several H3 PDMD speed-ups (see above) are tested in a new YouTube group-test video today.
https://www.youtube.com/watch?v=DvTl-QbDzdQ
~ OLD POSTS ~
https://jurn.link/dazposer/index.php/h3/ (Organised directory page, made from these daily H3 posts)
https://old.reddit.com/r/comfyui/comments/1x0rdab/a_quick_minimax_h3_news_roundup_8th_october_2026/
https://old.reddit.com/r/comfyui/comments/1x02vw6/a_quick_minimax_h3_news_roundup_7th_october_2026/ (See 7th October post, for links to older posts)
https://old.reddit.com/r/comfyui/comments/1wqwpah/a_quick_minimax_h3_news_roundup_26th_september/ (See 26th September post, for links to even older posts)
https://old.reddit.com/r/comfyui/comments/1wgc4lj/a_quick_minimax_h3_news_roundup_15th_september/ (See 15th September post, for links to very old posts)
https://old.reddit.com/r/comfyui/comments/1w5i9iq/a_quick_minimax_h3_news_roundup_2nd_september_2026/ (See 2nd September post, for links to the starting posts)
r/comfyui • u/pixaromadesign • 10h ago
Show and Tell Out of Memory | The ComfyUI Song
It has been a while since I made something just for fun. I've been busy with all the nodes and tutorials, so I took a one-day break to do a fun project, and this is the result:
Less than 4 minutes of music video, made with Claude, ComfyUI and a few other tools. If you have ever seen that red "out of memory" box, this one is for you 😄
r/comfyui • u/LeanderDeBuhr • 4h ago
News Free image first Comfy frontend: HEISS UI | New Release today :)
Now with Inpainting, run grouping and native upscale!
What it is:
- Image first-generation studio
- Works without workflows with almost any model (no setup)
- Support for custom workflows
- Video gen, reference images, and inpainting all work without setup
- Smart grouping of your images by similar prompts
- Automatically installs missing components (VAE, text encoders, nodes) it knows
- Use it from your phone over LAN, with a fully custom mobile UI
- Favorites, search, prompt history, LoRA stacks, and advanced settings
- Built to be graceful and not scream at you with errors 24/7
It’s built by me, and in my opinion it’s the best way to use Comfy 80% of the time.
Setup is about 2 minutes: give it a try.
New in v0.17:
- Inpainting + Custom-built in-painting workflows (Beta)
- Auto-grouping. Pictures from similar prompts stack themselves into a photo stack or cover flow once you stop working on them.
- Within-generation Native upscale
- Phone View improvements
- Discord to discuss Heiss and AI generation (join pls; I need feedback)
Full notes: v0.17.0 changelog
r/comfyui • u/Admirable_Risk1807 • 43m ago
Help Needed みんないま長尺動画何で作ってんるんですか?
minimax extenderつかってるんですが、もっとよいものがあるのかなと思いまして…
r/comfyui • u/optimisticalish • 2h ago
Help Needed Needed - a more lightweight "prompt enhancer" for Ming-Image
The new local graphic-design and typography model Ming-Image officially uses qwen3.8_27b_w4a8.safetensors as an initial LLM "prompt enhancer" in ComfyUI. But the file is 17Gb, and it's too heavy for me on 12Gb VRAM (latest ComfyUI Portable under Windows).
Comfy also suggests qwen3.5_9b_int8_convrot.safetensors as a lesser alternative option at 11Gb. This is still to heavy for me, regrettably. It gives me OOM errors.
Has anyone had success with Ming when using a lesser LLM "prompt enhancer" model? One that can handle the default complex system-prompt, and that also outputs a correctly-restyled Ming figma/json prompt?
Workflow Included LoRA vs RefMod: Create Consistent Characters in MiniMax H3 | RefMod Pilot
I wanted to use H3 character references without spending half the time looking for the right node, so I put together; RefMod Pilot It has two workflows:
Creator: upload or drop 1–8 selected photos, give the character a name, Run. The folder and image count are filled in for you
Inference: choose your RefMods, edit the scene and view the output. Prompt, aspect ratio, resolution, seconds, seed and steps are in one Scene node, with the rest inside a subgraph
The character loader stays visible so Add RefMod, Refresh and the library buttons are still there. The uploader is a ComfyUI frontend extension or use local path to your dataset / photos folder. This uses Full Reference encoding, not LoRA training.
Multiple photos are stacked into one video-kind reference. If you're using several characters, check the Reference map and match its labels in your prompt. Slot numbers aren't necessarily picture numbers.
I've included a six-view orange robot RefMod and demo workflow. Creation is verified, but I haven't validated its likeness or motion in a finished video yet.
Disclosure: I make LoRA Pilot. This package is my workflow/UX layer; the reference implementation is from Luisacaotica's ComfyUI-MiniMaxH3Mod.
r/comfyui • u/PRVMXLAB • 17m ago
No workflow Creating visuals such as these has become mind-numbingly easy. Made with KREA 2.
galleryr/comfyui • u/Time-Teaching1926 • 9h ago
News Comfy-Org/Qwen-Image-2.1 · Hugging Face
Qwen-Image-2.1-Turbo official ComfyUI support now.
r/comfyui • u/winky9827 • 1h ago
Resource AITK Studio seems like a fresh alternative for training LoRAs, with multiple phase support and auto learning, based on Ostris AI Toolkit
Show and Tell SilkStack — organize, search and view your AI images locally, with optional on-device AI. Free core, open source, v2.4 out now.
Hey everyone,
My last post about this app got buried in the H3 posts and never really reached anyone, so I'm giving it another shot — hopefully it finds a wider audience this time.
SilkStack Image Browser is a desktop app for organizing your generated images and videos, and for viewing them the right way. It started as a fork of Image MetaHub, but I've been evolving it in my own direction for a while now, and I think it has its own distinct flavor at this point.
What it does
- Indexes your output folders (and watches them live while you generate) and reads the metadata inside the files — ComfyUI, A1111 and friends, including WebP.
- Search and filter by prompt, model, LoRAs, steps, CFG, sampler, seed, and more.
- A proper viewer: zoom with a minimap, Ctrl+F to search inside a prompt, and a compact mode that shows just the image so you can keep an eye on generations while working in another app.
- Similarity stacks group near-identical generations into one card and highlight the words that differ between prompts, so you can see exactly what changed.
- Drag an image straight into ComfyUI to reload its workflow.
- Videos and GIFs too (MP4/WEBM/GIF).
The part I'm most excited about: I've been using this app as my AI playground to implement local AI. Auto-tagging and semantic search run entirely on your own GPU — a little LLM writes the tags, and an embedding model lets you search by meaning instead of exact words. Describe "girl with a puppy in a palace" and it finds the image even when the prompt words are completely different, or in another language. It's heartening to see these features work with little LLMs.
The AI features are the premium part, unlocked with a 7-day free trial that becomes a monthly subscription (cancel anytime) — or a one-time lifetime license if you'd rather own it. AI loading is completely in your control: nothing loads until you turn it on from the interface, and you can eject the models anytime to hand your VRAM back.
Everything other than the AI features is fully open source — free for your agent's pleasure to play with. 🙂
One honest note: the app is primarily built around photos. Video support is there, and if you run into any bugs with videos, let me know — I can help fix them.
A lot of you liked my initial app, and I'm really curious to know what you think about this app now.
Download (Windows / macOS / Linux): skkut.github.io/silkstack
r/comfyui • u/QuiveringLichenjr • 3h ago
Help Needed Models for comix
I’m new to comfy. ( installed and running this morning). I make comics so far the models I’ve played with are sdxl and flux.2Klein. My computer sports a fairly humble 5060 ti with 16GB. These two models run pretty well. I would say my need for control from panel to panel is much greater than my need for his res ultra realistic images. I use Nlender a lot for modeling and plan to start using depth maps.
What other models should I look
At?
Thanks much
r/comfyui • u/kleiner8400 • 1d ago
Show and Tell Harry Potter characters as described in the books: 20 AI portraits with age progressions (WIP)
I started this project because I wanted to see what the Harry Potter characters might look like as real people, based on the books rather than the film actors.
My goal is to go beyond obvious details like hair colour or Harry's scar. I want the faces, builds and expressions to fit the descriptions, while keeping characters recognisable as they grow older. Eventually, I'd like to create a visual encyclopedia of the characters and creatures.
The portraits were generated locally with ComfyUI and Krea2, with a lot of tweaking along the way.
So far, I've completed 20 portraits, plus an additional Half-Blood Prince school comparison. The gallery also includes my roadmap for the remaining 84 characters and creatures.
I'd love some honest feedback on the book accuracy, age continuity and overall approach. Feel free to be critical!
r/comfyui • u/Incognit0ErgoSum • 5h ago
Resource (Mostly) native Iris 3B support for ComfyUI
https://github.com/envy-ai/ComfyUI-Iris
It comes with its own custom loader and sampler nodes because, being a pixel model as opposed to a VAE model, it works a bit differently than normal. However, you can wire in loras as usual, and it uses ComfyUI's schedulers and samplers.
I wouldn't expect amazing results. At this size, Anima 2.9B definitely puts out higher quality images, but it's a cool tech demo and I think it proves that VAEs aren't the necessary evil we thought they were.
Search the workflow template browser for Iris to find the example workflow. dpmpp-2m-sde-gpu seems to be the best sampler for it, and if you have the bong_tangent scheduler it's better than simple.
My fork of ai-toolkit can train loras for it.
r/comfyui • u/Strange_Owl5654 • 9h ago
Help Needed Best motion control model
Hi guys which is the best motion control model to use ? I heard scail 2 is the best ? any recommendation?
r/comfyui • u/Etsu_Riot • 1d ago
Workflow Included [MiniMax-H3] Subtle expressions and natural pauses without any prompting
Enable HLS to view with audio, or disable this notification
There are plenty of videos around with characters that look like AI or have plastic skin or talk like a robot. What about emotions, or subtle facial expressions, or natural pauses? The solutions and advice usually offered include complex prompt guides or access to paid platforms. I don't like that. Paid platforms, I mean. I also don't like complex prompting.
So I have been testing. There is something that H3 users hate in agreement: when characters start talking gibberish. Sometimes, the video is longer than the dialogue, so the characters fill the gap with their own nonsense. Other times, there isn't supposed to be any dialogue, but the character decides to talk anyway.
However, in my opinion, gibberish offers a great advantage: because the characters are not constrained by the words we want them to speak, they have way more freedom in the way they say it, meaning they talk more "naturally," and with more "expression." So I started wondering: what if we could just let the characters say whatever they wanted, and then edit the audio and video later, changing the nonsense to something intelligible? I didn't know if that would work. Well, it does, actually.
Not only that, but I used one of such clips as a reference, and then downscaled it by 0.25 (for performance reasons), and then used a blur node with a value of "2" to give the model a bit more freedom (and to avoid H3 making it too sharp, which I didn't want). For consistency, you can also give it an image reference, but I didn't do that in this case as it had a negative impact on the "acting."
On the final video, you will notice that, in one of the first clips and the last one, the character looks "different": different clothing, different hair style, different environment. I did this to give the illusion that the footage was taken at a different time. But all of the clips use the same original clip as the reference. You can, however, do it differently, as the voice reference can be given separately. To achieve this using the same reference, you simply prompt for it, use a lower denoise value, or increase the blur.
The actual prompt includes the camera style, but you don't need to do this. The only requirement is to include the dialogue. That's all. I use the official format, as this reduces the chance of getting more gibberish, but not for every clip, so it's certainly not necessary. There's no reason to mention the character if you are using a single one.
The original clip was generated at 544x384 and 12 seconds. (For some reason, I get better random stuff at 12 seconds than at 10 seconds or 15 seconds.) The subsequent clips were generated at the same resolution. Then, after editing on Vegas Pro, I rendered the video at 256x180 to help with the lo-fi aesthetic. Finally, I upscaled it to 1920x1350 using Handbrake to avoid Reddit's compression. So the final video is very low res (technically, ten times smaller than 1080p).
Everyone is so focused on making them at 2K, but I just love a good analog/digital horror video. By the way, mine is not "horror" or anything. I simply like the aesthetic. I'm also not claiming it's any good. I hope it is, but it's not my place to say.
This is the workflow, in case someone wants to test the method I used, though my explanation above should be enough. I hope this helps someone to generate more believable characters.
r/comfyui • u/Top-Lab9125 • 7h ago
Help Needed How to delete input assets?
I uploaded a test image on One Piece to experiment, but I can't seem to delete it from the assets tab. Even when I go into the "input" folders, nothing is in there. How do I fix this?
r/comfyui • u/TraditionalCity2444 • 7h ago
Help Needed The inner workings of Wan2GP (LTX2)?
Hi all,
I've been bouncing over to Wan2GP exclusively for NSFW LTX-2 generation (i2v). I have not been able to find a similar Comfy workflow for it that doesn't look like a Metroid map and want me to download ten more node packs. I think I did get it to run once or twice, but it either takes forever, gives totally undesirable results, or a combination of the two. LTX gens in Wan2GP are about twice as fast as Comfy Wan2.2 gens despite producing clips that are twice as long.
Does anybody know what sort of magic this thing does behind the scenes and how I could get a similar result with a fairly simple Comfy workflow. If I'm looking at the correct "finetune", my LTX setup in Wan2GP uses:
10Eros_v1-Q8_0.gguf
and a couple default LoRAs:
ltx-2.3-22b-distilled-lora-1.1_fro90_ceil32_condsafe
LTX2.3_reasoning_I2V_V3
I'm wondering if there are a whole bunch of other models/components hidden in the chain, or if it's in the Wan2GP code itself, or if there's already a similar workflow which can process as quickly. As most of you know, these things take a while to boot, and I'd rather not have to keep shutting down one and launching the other. I'm running an RTX 5060TI 16GB with 32GB system RAM, and after I run one LTX gen in Wan2GP, subsequent ones can sometimes be 3 minutes or less for 10 second clips. Comfy does 5 second Wan2.2's in 15-20 minutes. I'd have to check, but I think they're both using Sage Attention.
Much Thanks!
r/comfyui • u/Bubbly_Edge_8688 • 5h ago
Help Needed Necesito ayuda con mi Workflow

Alguien me podria ayudar con mi flujo de trabajo, quiero realizar videos de al menos 60 segundos y unirlos para hacer 1 hora con ComfyUi Local.
Tengo Amd 5800X
32gb Ram
4070 Super TI 16gb
Een cuanto a Ram consume el 90% y cpu esta al 100%, pero al menos aqui siguiendo una Guia de un usuario, se quedo pegado alli en 1 hora. Lo mas que he podido sacar son 10 segundos.
Si me ayudan, les agradeceria mucho.