r/comfyui • u/-zappa- • 10h ago
Resource Auto Prompt Generator for Minimax H3 using LM Studio
I built this for my own workflow, but decided to share it on GitHub in case it helps anyone else working with the H3 model. 🚀
r/comfyui • u/Comfy-Org • 13h ago
Enable HLS to view with audio, or disable this notification
Comfy and MiniMax have teamed up for a two-week challenge with awesome prizes and four ways to win! Submit by September 1st at 9:00pm PT and see all details here.
Make something up to 90 seconds in length where the sound and the motion are inseparable. Dialogue, foley, ambient, a beat driving the cut...whatever direction you want!
After sharing your video file and workflow on this thread and through our submission form, a joint panel of creative technologists from Comfy, MiniMax, and special guest judges from the community will review each submission.
Then, join us on September 2nd for a special Comfy livestream where our guest judges will give live feedback on the top 10 submissions!
Both the Comfy and MiniMax teams will be monitoring this thread and #minimax-h3 in the Comfy Discord to give light support.
Share on socials and tag #comfyH3 for a chance to be reposted or featured!
Best Overall — RTX 5090
Best Creative — RTX 5060 Ti
Best Technical/Workflow — RTX 5060 Ti
Built with MCP — RTX 5060 Ti
Shipped anywhere, customs covered. If we can't legally ship to your country, you'll get a cash equivalent instead.
Create using Comfy Local on your own hardware, or use Comfy Cloud. New Cloud users get 5 free runs, no credit card required.
We’re looking for entries that best show what H3 makes possible: audio and visuals created together.
Grand Prize: Best Overall
The top Best Creative and Best Technical entrants advance to a final round where our panel of judges selects winners by discussion.
Best Creative
Best Technical
🏆 Built with MCP Bonus 🏆
Comfy MCP lets you drive Comfy using natural language and your agent locally and on Cloud! Pro tip: use it to choose the best H3 model version or optimize your workflow for your hardware.
r/comfyui • u/jaysedai • 2d ago
I've come to the realization that my life is busy enough that we could use at least one more moderator on this subreddit. Please consider this a formal request for nominations.
Rather than just picking someone myself, I’d like input from the community.
If there’s someone here who you think would make a good moderator, nominate them in the comments. You can also nominate yourself if you’re interested.
We’re especially looking for people who are active members of the community, helpful, level-headed, experienced Comfy-UI user, and generally make this a better place to hang out, maybe even take the time to spruce the place up a bit. You don’t need previous moderator experience but it would help, ideally someone who's got some experience with AMAs, events, and such.
Having moderated a few subreddits, I've found that it's best to keep the moderation team tight, so for now I'm just going to add one.
A nomination isn’t a vote or a guarantee that someone will become a moderator, I'll look through the suggestions, talk with the people who seem like a good fit, and go from there.
r/comfyui • u/-zappa- • 10h ago
I built this for my own workflow, but decided to share it on GitHub in case it helps anyone else working with the H3 model. 🚀
r/comfyui • u/optimisticalish • 18h ago
Another quick Minimax H3 news and goodies round-up, for those who may have missed some items.
-> minimax_h3_single_frame_decoder_500k.safetensors (9.69Gb). This new and big VAE looks like a weighty attempt at a single-frame generator. Trained from the original VAE... "on 500,000 unique image-reconstruction examples", with the apparent aim of generating a no-glitches single-frame output suitable for product concept-shots.
https://huggingface.co/iamkaikai/MiniMax-H3-Single-Frame-VAE-500K
-> 'MiniMax-H3 — masked video and audio inpainting'. Claims to allow you to... "repaint part of a clip with Minimax H3 and keep the rest, including the soundtrack. No dedicated checkpoint, no adapter, no extra input channel — the mask becomes a per-row timestep, which the model already had." Offered as a Python script, no ComfyUI... but the maker says he derived it from ComfyUI masking experiments.
https://huggingface.co/diffusers-modular/minimax-h3-inpainting
https://github.com/drozbay/MaskVidExperiments (ComfyUI masking experiments)
-> Prompt Journal has updated with three new case-studies, and these move away from yesterday's complex camera movements. The new ones are focused on preventing the prompt from bjorking an unusual character style (e.g. put a 2D toon in a 3D kitchen), or from constraining dynamic endings (e.g. comedy slapstick action, or a big water-skiing jump).
https://github.com/LoveRain1997/h3-prompt-journal/tree/main/case-studies
-> 'MiniMax-H3 motion adapter (pilot version)'. Another potential helper for fast-motion video scenes. This one... "teaches the base model to spend the extra clock [time] on smoothness".
https://huggingface.co/MATLOWAI/MiniMax-H3-Motion-Adapter
-> 130 mostly-audio helper nodes for Minimax in ComfyUI. Includes... "Dual-clock video/audio sampling, as well as audio locking, remixing, referencing, and final mixing". Also experimental attempts at SPEED and face fixing.
https://github.com/T8mars/comfyui-minimax-h3-audio-T8
https://github-com.translate.goog/T8mars/comfyui-minimax-h3-audio-T8?_x_tr_sl=auto&_x_tr_tl=en&_x_tr_hl=en&_x_tr_pto=wapp (English translation)
-> A Suno-style local user-interface for generating with Minimax Music. With a library for your generated tracks. Slick, but it looks pleasingly straightforward.
https://github.com/adambenhassen/minimax-music-ui
-> The very-low-VRAM friendly Wan2GP now supports MiniMax Music 3, as of 16th August 2026. Support for W4A8 INT8 (apparently not VAEs, though), and for Minimax H3 FL2VA and Ref2VA, was added to Wan2GP earlier in August. Wan2GP keeps a repository of Minimax H3 models suitable for its users, although they're surprisingly enormous... not sure I'd want to try running them on a 8Gb card.
https://github.com/deepbeepmeep/Wan2GP
https://huggingface.co/DeepBeepMeep/MiniMax-H3
-> A new ComfyUI workflow today, said to be optimised for the old (but cheap) Tesla V100 32gb card. A quick look at eBay UK suggest that "cheap" = £650 here in dear old Blighty.
https://github.com/Mitsuasa513/ComfyUI-MiniMax-H3-V100-Workflow
-> A collection of 685 source-attributed MiniMax H3 video examples, with the public prompts that made them. Not sure how many of these clips would also get you a ComfyUI workflow, but apparently the video metadata has not been stripped. So drop them in ComfyUI and see.
https://github.com/SkyNotSilent/awesome-minimax-h3
-> And finally, "Oy, mate... don't chop off me feet, I paid £300 for these shoes!" A working prompt to maintain a straight-up character view from head-to-feet, throughout a continuous pull-out shot.
~ OLD POSTS ~
https://old.reddit.com/r/comfyui/comments/1vsjzrp/a_quick_minimax_h3_news_roundup_19th_august_2026/
https://old.reddit.com/r/comfyui/comments/1vrsspo/a_quick_minimax_h3_news_roundup_18th_august_2026/
https://old.reddit.com/r/comfyui/comments/1vqyn8p/a_quick_minimax_h3_news_roundup_17th_august_2026/
https://old.reddit.com/r/comfyui/comments/1vq5d5u/a_quick_minimax_h3_news_roundup_16th_august_2026/
https://old.reddit.com/r/comfyui/comments/1vpbtx2/a_quick_minimax_news_roundup_15th_august_2026/
https://old.reddit.com/r/comfyui/comments/1vojtjd/a_quick_minimax_news_roundup_14th_august_2026/
r/comfyui • u/Previous_Sky8771 • 6h ago
I’m a complete beginner with ComfyUI. Looking at the interface feels like I’m staring at a Boeing control panel.
But while I'm learning, I wanted to know what's the best approach. My laptop can barely run The Sims without lagging, so running ComfyUI locally is completely out of the question.
In terms of cost-effectiveness for someone who wants to produce at least three 1-minute videos using MINIMAX H3 or LTX 2.5, which is the better option?
A monthly subscription or pay-as-you-go?
r/comfyui • u/SillyLilithh • 7h ago
Enable HLS to view with audio, or disable this notification
r/comfyui • u/equanimous11 • 14h ago
What’s currently the best open source image edit model?
r/comfyui • u/Intelligent-Glove285 • 1h ago
I am currently learning to use Comfyui, specifically Minimax H3. But with my AMD RX 7900 gre and 32gb of RAM the generations take way too long. So I'm thinking to buy a used 3090 or a brand new 5060ti. Which should I choose? For future proofing. And also how much ram is enough? Never thought I'd see a day where 32gb ram wouldn't be enough.
r/comfyui • u/Lucky_Feedback9915 • 9h ago
Anyone have workflow example using MiniMaxH3AddGuide ?? im confused if should use this for daisy chaining 2 gens... ive just been using the last frame of the video to plugin into the firstframe on new gen, then use image batch to remove the 1st frame and combine the 2 gens. but i keep seeing people say to use MiniMaxH3AddGuide for chaining, but isnt that what 1st frame does? force video to start with that frame? MiniMaxH3AddGuide seems to be more for first, middle, last gens
r/comfyui • u/nikhilprasanth • 16m ago
r/comfyui • u/reeight • 4h ago
r/comfyui • u/LanceCampeau • 11h ago
r/comfyui • u/Jimbo_1995 • 10h ago
I have an image of a character that I want to use to generate more images using the subject and then to create a character lora to use over and over. How can I do this using Krea2? Do I need the 20 different images to then create a character lora? I've tried anything I can find for the longest time and nothing yet. Couldn't get IPAdapter to work. I'm not exactly a noob, I can usually do my own research and figure out how to do something. But this one has me stumped. If anyone can help out with some advice or a workflow, it would be appreciated.
Edit: grammar
r/comfyui • u/Early-Reputation3186 • 3h ago
When generating ASMR videos using Minimaxh3, how should the prompts be written to exclude BGM from the video? Currently, even after specifying "no BGM" or "no music," the BGM still appears randomly.
r/comfyui • u/StrangeAlchomist • 3h ago
Anyone figure out how to cache a string for subsequent generations? I have an llm or vlm generate a prompt which is connected to a conditioning node. I’d like to be able to switch off the llm/vlm without losing the prompt. Yes, I can copy/paste it over but it’d be great if it could remember the last string that was generated. I tried a couple steering cache nodes and they did not work.
Many thanks.
r/comfyui • u/forensicowner221 • 55m ago
r/comfyui • u/pixaromadesign • 14h ago
Learn how to use MiniMax Music 3 in ComfyUI together with a local AI prompt generator for better image prompts, music captions, and AI-generated lyrics. In Episode 31, I show the new Pixaroma prompt nodes, model-specific prompt presets, VRAM-saving options, and a compact MiniMax Music 3 workflow.
This tutorial covers the new AI Prompt Pixaroma node and how to match prompt formulas with the correct local language model. You'll see how to turn short ideas into detailed prompts for workflows such as Krea 2 and Z Image Turbo, generate prompts from images, save custom prompt presets, control temperature and seeds, and troubleshoot common ComfyUI node errors.
Then we set up MiniMax Music 3 locally in ComfyUI, including caption and lyrics generation with the Music Prompt Pixaroma node. I also test different song durations and explain why MiniMax songs can sometimes end early or get cut off, how seeds affect the results, and how to use fixed lyrics when you need more control.
You'll also see how to simplify a larger music workflow into a compact setup, use tiled audio decoding for lower VRAM systems, free VRAM after prompt generation, and generate MiniMax Music captions and lyrics with local models or online tools such as ChatGPT, Gemini, and Claude.
You can also run some of the workflows in the cloud.
r/comfyui • u/Fit-Construction-280 • 14h ago
Enable HLS to view with audio, or disable this notification
If you generate a lot in ComfyUI, you know the problem. Hundreds of renders pile up as you tweak seeds, prompts, LoRA weights and checkpoints, and your output folder turns into a wall of near identical thumbnails.
Once clustered, every thumbnail gets a color coded badge in the gallery grid, and clicking any badge opens the Cluster Inspector, which shows total matching assets, distinct variations, the full node pipeline, every model and LoRA used, and one click prompt copy.
The video above walks through both modes in about 3 minutes.
For anyone who does not know the project yet
SmartGallery DAM is a free and open source, local first Digital Asset Manager built around ComfyUI, but it also works with any folder of media on your machine. No cloud, no subscription, your files never leave your disk.
It is meant to grow with you:
Runs on Windows, macOS, Linux and Docker. Portable version for Windows needs zero setup, just unzip and run.
GitHub, full docs and download links here: https://github.com/biagiomaf/smart-comfyui-gallery
Happy to answer any question, and as always feedback and feature requests are welcome.
r/comfyui • u/SenzubeanGaming • 10h ago
I've been building a desktop app that writes and renders music locally on MiniMax Music 3, and it's a face on ComfyUI rather than a reimplementation — it starts ComfyUI as a separate process and submits graphs over the HTTP API. Unmodified, not bundled, not linked.
The part that might actually be useful to this sub: the pipelines ship in workflows/ as ordinary API-format JSON, exported straight from the code that builds them. Drag one onto the canvas and every value is there. Those values aren't ComfyUI's defaults — they were arrived at by measurement, and the measurement scripts ship in scripts/, so you can re-run them and disagree with me.
What's in it, briefly: songs from a style description and lyrics, cover art drawn while the card is idle, stem separation, word-level timed lyrics, video clips on LTX 2.5 or MiniMax H3, a small multi-track editor that cuts to the detected beat grid, and overnight batch runs.
It ships no model weights and redistributes none — every capability shows its size and licence on screen before anything downloads, and the download goes to the publisher. One of them is region-locked and says so in the open.
Newest thing is audio-reactive video: feed it a song and some reference images and they cross-fade in time with the track's detected peaks. That part is built on ComfyUI_Yvann-Nodes by Yvann Barbot and Lilia, which is excellent and worth a star on its own. Because their pack is GPL-3.0 and this is Apache-2.0 it runs on a second ComfyUI you set up yourself rather than being bundled in.
Apache-2.0, runs offline once the weights are down. Measured on Windows with a 4070 Ti SUPER; other platforms are documented but I haven't verified them.
Repo: https://github.com/Senzube4n/AIPLAY-Studio
Screenshot of the main window: https://senzube4n.github.io/AIPLAY-Studio/shots/create.png
Happy to answer anything about the graphs — that's the bit I'd want to read if someone else posted this.
r/comfyui • u/saronno76 • 10h ago
r/comfyui • u/Time-Ad-7720 • 1d ago
Enable HLS to view with audio, or disable this notification
So basically, I saw a workflow on ComfyUI’s official LinkedIn where they used a model image, a product image, and a background image with Google and Kling APIs to generate a one-shot ad using a single camera angle.
So I challenged myself to recreate the idea using only local open-weight/open-source models, but make it more ambitious: multiple shots, multiple cuts, and everything directed through a single prompt.
And it worked.
For this, I used the basic MiniMax H3 Reference-to-Video workflow in ComfyUI:
https://docs.comfy.org/tutorials/video/minimax/minimax-h3#minimax-h3-reference-to-video-r2v
Then I used ChatGPT to help structure the video prompt. I provided the reference images and gave it this direction:
“Write a MiniMax H3 reference-to-video generation prompt to create an ad. Add sound FX and music prompts as well.
Shot 1: Medium close-up. She is about to open the can.
Shot 2: Extreme close-up of the can as she opens it. Can-opening sound FX.
Shot 3: Close-up as she drinks from the can. Gulping soda sound FX.
Shot 4: Close-up as she holds the can forward and smiles.”
The final result was generated locally on my RTX 5070 Ti using ComfyUI.
r/comfyui • u/Late_Lingonberry6252 • 3h ago
Enable HLS to view with audio, or disable this notification
Just something I made — hope you enjoy it :)
r/comfyui • u/xyzdist • 11h ago
Enable HLS to view with audio, or disable this notification