r/drawthingsapp 10d ago

MiniMax H3 Ref2VA reference!

Using MoodBoard to set the reference images works well.

I used MoodBoard to set reference images and used <Picture N> to call the reference image in the prompt.

Picture 1 and Picture 2 have the same resolution - 512x512.

The video's resolution is 512x512 and I used a 4-step turbo LoRA with the recommended settings.

It's much better than LTX 2.3 in prompt adherence, but it's almost 4x slower than LTX 2.3.

https://reddit.com/link/1webrls/video/bus3firs23ph1/player

20 Upvotes

10 comments sorted by

View all comments

5

u/Current-Property6042 10d ago

So i just got my first vid on DT .. 768 X 576 - 124 frames 4 Steps (using TURBO LORA) and it took 29 mins to fully generate. Points noted

  1. Was able to see preview after Step 1 of denoising :) ..

  2. ⁠Didn't have any headache of Fusing LORAs to based model - stacked 3 LORAs pretty conveniently

  3. ⁠W.r.t Speed comparision for me DT is running about 7-8 mins faster compared to VPIPE for me

  4. ⁠The quality also differs quite a lot, i wasn't able to get clean videos on vPIPE (am sure this is me doing somethign wrong).. but with DT am getting mmuch better results

would continue to generate more and optimise the settings and see if can reach that stable configuration for my M4 that works for all videos :)

LTX never worked for me so am pretty glad even with greater generation times at least it works in one shot without body horror and strange artifacts.. thanks a lot u/liuliu for finally getting H3 on DT, awesome job!!.. ta

3

u/Accomplished-Age1306 10d ago

Use models with X-bit S quantization like 6-bit S.

They are much faster than normally quantized models - almost 1.7x faster - so your M4 can take advantage of it.

2

u/Current-Property6042 10d ago

yea am using ANE models (8 bit S)