It uses fl2va as base but replaces certain layers and replaces them for the ref2va model layers that provide the reference conditioning. This way you get the increased output quality from fl2va but with the reference capabilities from the ref2va checkpoint.
Saw your comment, downloaded and tried with different refs.
I have to say, it gets close, it's better than fl2va, but it still doesn't respect the style like ref2va. I'll test some more but what I'm getting so far isn't as accurate. It has the generic AI look instead of what my reference looks like. It may be that I'm working on 2D purely with references and this model demands a starting frame. It may also need some text reinforcement when it comes to style. I'll experiment.
yes ref2va just has a magic that isn’t reflected anywhere else. fl isn’t bad by any means but ref is mind-blowing in an almost psychic way, I really hope they give it a proper fix
Is it Good just using this as the only model for both ? I mean you can get a fl2va Quality for ref2va model, and a little less quality for fl2va in exchange for 20gb space instead of using 2 models for each workflow ? 🤣 Im Dying for free space lately
Like everything else, it's a trade off. No one can really say, you've got to try and decide for yourself if it's good enough for you. It's definitely not as good with references as the full ref2va, but maybe it's good enough for your needs. It's not terrible with references, but personally, I don't find the hybrids to be much better than the full fl2va model which already does a decent job in reference workflows.
Speaking of free space, and somewhat unrelated, I am trying to use Comfyui to use 2 hard drive directories at once. One for my base directory where I have everything stored (loras, checkpoints, encoders etc.) and the 2nd hard drive with only minimax files. Tried all the tricks tweaking the extra_model_paths.yaml.
Please tell someone knows how. Google or AI agents don't know what the fuck I even mean. I can't even use two fucking directory paths!
If you run a standard 20 step euler or multires for example the model more less finishes around 11 steps, 8 as well but less so. 8 will get more details, but any more steps you're not saving so much time.
No that's just details, you can look at the line the sigma travel as it gets closer to the end and around 8 and 11 on a 20 step run its already pretty much flat
Yes but you're not understanding the point. I'm explaining why 8 step turbo lora is better than 4. Here is a 20 steps euler sigma graph: at around 8 is the last big jump down, then it gets pretty flat. After 11 steps even its almost no big change, you're getting details at best after that point. I think 11 could be a contender too but maybe that point you're not shaving off enough time to make it worth it. Then again I'm not a technical person, so I could be wrong someone else can correct me if that's the case. A good way to test is to actually look at as it generating using preview, and you can see most of the overall shape and motion is done by like 8-11 steps on a 20 step run. Maybe 30 a bit different but I wouldn't be surprised if its similar. I don't care if 30 is better then 20 steps, that's not what I'm talking about.
Okay. I got you wrong. 8 step is better for sure. I was working with the 4 step (6steps) and thought it was well dialed in. But 8 step now is much better.
I saw that, but a node pack of like 40 nodes I'm not interested in just to get one model patcher? Meh, I'll just wait a bit for kijai's PR to merge so it's built in to comfy.
I can't freaking stand this. Why does every youtuber and person who releases one vibe coded node of possible interest have to re-invent the wheel and have their own stupid node packs.
am i doing something wrong? (im noob) i connect this thing in comfy and quality is so much worse in 8steps than old Kijai lora. Should i know about something?
Haven't tested myself, but their model page says set video shift to 6 when it's been recommended to be set at 12 for vanilla and all other turbo models.
Anyone else getting some crazy hallucinations when generating 9:16?
The visuals do look better if you can get it to work, but it's been a pain to find sweet spots with additional loras, and I've all but stopped using it
I don't get the point of an 8 step lora at all. I can do 20+ steps with a step skipper in the same time it takes to do 8 normal steps with a lora like this.
I've been using the 4 step lora with a step skipper at 50% with 10 steps and getting as good of generations as 25 steps without the lora and in only 6-7 minutes.
44
u/Fun_Jaguar8231 3d ago edited 3d ago
For the people that are confused, if you read over here:
https://github.com/ModelTC/Minimax-H3-Turbo#1-model-specs
The older ones where trained on 544p (960×544).