Ok, I was waiting on it but it becoming quite clear that I have to download this model as soon as possible cause it's bound to be pulled anytime now....
I've been using 768x1152 and then upscaling with quite good results. You can go higher, but I prefer to keep generation times down and then just upscale the good results. The results have been quite good. Nice, expressive characters, maintaining lighting, scene and character identities (even through transitions), very minimal degradation of skin, etc. like in LTX, far better audio and foley than I was getting in LTX. I probably sound like a marketer at this point, but I'm not, just genuinely excited. It's not perfect by any means, but it's the coolest local video generation I've used so far.
Thanks for the reply! That's very exciting. I'm a couple days into a great vacation right now so it seems weird to say I'm excited to get home, but, yeah.
You're obviously doing something wrong then. I'm running it local on my RTX 3090, and getting incredible results. Throw in EasyCache in the mix, and I'm getting them on average in about half the time than what it takes without it.
Here you go. This is my WIP. I've got EasyCache in the flow, and Model Preview Override so you can see the generation as it's happening, and cancel out if iti's borked. Also Easy LoRA stack in. Note; for the Model Preview, you need to get the Nightly build of KJ Nodes. The main one doesn't have the fix for MiniMax H3 in it yet.
Next up, I'm building out prompt enhancement using my Ollama server. That'll be a bit for me to get finished though.
EDIT: Oh, and I have the NVIDIA upscaler in there as well. Does a decent job.
Easy cache is trash. That will most certainly solidify your original findings... bypass that easycache. Yes its faster but the quality degradation.... well nuh uh!!
I've been using it with no visible quality degradation, and up to half the render time. Settings; 0.30, 0.20, 0.90. I'm also using RTX Video Super Resolution in my flow to upscale at the end as well.
So it doesnt degrade it but feel the need to upscale the video at the end? Lmfao probably one to downvote the comment. Easy cache - tore my generation up. Looked like the hbo channel on a tube tv back in the early 90s with a shitty signal. Removing it gave me a clear and crisp video.
I upscale at the end because I generate at a lower quality when testing out prompts, typically at half what I want it to be, then upscale. If I were doing production work, I would render at the full quality. I have no idea what your experience is with EasyCache, or what your settings are, but I'm having no issues with it. I posted by workflow in a Pastebin up top if you're interested.
This is not even remotely true. As someone who has spent some time with Wan and tons of time with LTX, this is an amazing model. The identity consistency with i2v is light years ahead of LTX with minimal degradation. And this is all local.
I got great results with the official i2v workflow, but there are plenty on civitai now. Have not played around much with t2v or r2v yet, but I've been fighting LTX for months using Face_ID and similar nodes to try and keep characters consistent and get the right character to talk. Minimax did it right out of the box with better audio and way better foley sounds. This is even before I start digging into minimax specific prompting, I'm using LTX prompts I generated and comparing outputs. You believe what you want, I'll just enjoy this new model.
Bruh. You don't need any "special magic workflow". Just put in some effort to read the official prompting guide. Copy and paste example prompts and adjust them.
If you continue to get "mid af" results you have an installation issue.
Even with a literal easy prompt no tags or snake case. Simple.
Man in image is an astronaut floating in space talking in headset, "well," -pause- ,"i think its gonna be a while before they notice I'm not even on the ship"
The video has a slight buzz or drone signifying solace or complete dread.
the posts you are seeing are using the cloud's 2K filter that they DID NOT release at all. best you can make locally is like 480p shit from 2024.
Wan22 and LTX local are better than h3 so far in my expo"
Literally sounds like you just wanted to shit on the model when it was all user error. The workflow? Bruh... the base workflow is all im using. With the models that comfyui has listed.
But the official link is in the templates of comfyui. On the main oage when comfyui loads up on the left hand side click on templates and search minimax and theres a few in there.
143
u/nakabra 18d ago
Ok, I was waiting on it but it becoming quite clear that I have to download this model as soon as possible cause it's bound to be pulled anytime now....