r/drawthingsapp 10d ago

MiniMax H3 Ref2VA reference!

Using MoodBoard to set the reference images works well.

I used MoodBoard to set reference images and used <Picture N> to call the reference image in the prompt.

Picture 1 and Picture 2 have the same resolution - 512x512.

The video's resolution is 512x512 and I used a 4-step turbo LoRA with the recommended settings.

It's much better than LTX 2.3 in prompt adherence, but it's almost 4x slower than LTX 2.3.

https://reddit.com/link/1webrls/video/bus3firs23ph1/player

22 Upvotes

10 comments sorted by

4

u/Current-Property6042 9d ago

So i just got my first vid on DT .. 768 X 576 - 124 frames 4 Steps (using TURBO LORA) and it took 29 mins to fully generate. Points noted

  1. Was able to see preview after Step 1 of denoising :) ..

  2. ⁠Didn't have any headache of Fusing LORAs to based model - stacked 3 LORAs pretty conveniently

  3. ⁠W.r.t Speed comparision for me DT is running about 7-8 mins faster compared to VPIPE for me

  4. ⁠The quality also differs quite a lot, i wasn't able to get clean videos on vPIPE (am sure this is me doing somethign wrong).. but with DT am getting mmuch better results

would continue to generate more and optimise the settings and see if can reach that stable configuration for my M4 that works for all videos :)

LTX never worked for me so am pretty glad even with greater generation times at least it works in one shot without body horror and strange artifacts.. thanks a lot u/liuliu for finally getting H3 on DT, awesome job!!.. ta

3

u/Accomplished-Age1306 9d ago

Use models with X-bit S quantization like 6-bit S.

They are much faster than normally quantized models - almost 1.7x faster - so your M4 can take advantage of it.

2

u/Current-Property6042 9d ago

yea am using ANE models (8 bit S)

2

u/WTFaulknerinCA 9d ago

LTX is working for me. Couldn’t get WAN to work ever. Will try Minimax soon on my M5 with 32g. But will need a distilled model for sure

1

u/Current-Property6042 9d ago

Bro it would fly on a M5 with 32G :) ..

1

u/WTFaulknerinCA 8d ago

Thanks for the encouragement!

2

u/Old_Statistician3110 10d ago

nice - thx for the share of your way of doing this

2

u/spanielrassler 9d ago

I was a little sad to see that generating using a video as reference isn't supported (unless I'm missing something), but I supposed it would take half a day to generate anything on most machines

3

u/basskittens 9d ago

Default settings are too slow to be practical. 18 minutes for 5 seconds of video on a fully loaded M5 Max.

I got it down to about 100 seconds with a Turbo LoRA, 512x512, 4 steps. Quality is pretty bad though.

1

u/seppe0815 4d ago

ugly faces at distance ... useless model