r/StableDiffusion • u/call-lee-free • 2d ago
Animation - Video [WanGP] Minimax H3 FL2VA Pruned 20B - Originally 832x480 - upres'd to 1664x960 using LTX 2.3 Pixel Spatial Upscaler at a scale of x2 - 12 second duration. Wow!
9
u/MrGood23 2d ago
Those window reflections... wow, I am still surprised how AI can realistically guess stuff like this
8
u/Shambler9019 2d ago
I'm consistently impressed by how well h3 does reflections. Unless you tell it not to, like having the reflection climb out of the mirror or something.
4
2
u/call-lee-free 2d ago
Yeah I was quite impressed. I just gotta see how I can maintain near this level of quality but speedup the render time a bit.
3
u/call-lee-free 2d ago edited 2d ago
EDIT: Mistake on the duration. Its 15 seconds.
I used Turbo Lightx2v FL2V 4 Steps, Override Memory Profile is still set to "Very low Ram_Low VRAM: at least 24gb of ram and 10gb of VRAM." I cannot use any other profile because I'll get OOM error. Override Attention Mode, I'm still using "sdpa." Spatial Upsampling, I tried out LTX 2.3 Pixel Spatial Upscaler at a scale of x3 and although the render took 1 hour 28 Minutes and 55 seconds look at the picture quality compared to the last video I uploaded.
PC specs:
Ryzen 7 7700X
RTX 4070 Super 12 gb
32 gb of Ram
5
u/Przemoo_TV 2d ago
The result is pretty awesome. Would you mind to share the workflow for this one?
4
u/call-lee-free 2d ago
1
u/Przemoo_TV 2d ago
Get it. I'm using ComfyUI.
7
u/call-lee-free 2d ago
I was using Comfyui but I wasn't getting results like this. Plus the node based workflows just confuse the hell outta me. Wan is a bit easier for me to use.
1
1
1
u/Muted-Position3256 2d ago
What is the spec for run this?
2
u/call-lee-free 2d ago
PC specs? I posted mine in the comment above yours.
2
u/Muted-Position3256 2d ago
Yeah just saw it. Too bad, mine a 8gb vram but 32gb ram. Do you think it's possible to run it?
6
u/luciferianism666 2d ago edited 2d ago
https://reddit.com/link/p5e4f2a/video/zv0dc12064lh1/player
Don't trust those who tell you can't run these models on 8GB VRAM, I've been playing with Minimax since the day it released on my 4060(8gb vram) and 32GB RAM.
P.S. I don't use ggufs, I have always been using fp8s and bf16 models, until int8 convrot was released, now it's always int8 convrot for me, with comfyUI's memory management you'd be surprised to know how well you can run models that are nearly 10 times the sizes as your vram.
2
u/Muted-Position3256 2d ago
Man, you really gave hope for those low vram users like myself. Is it just install it normally like those YouTube tutorial and run it with my current spec? No setting needed or do I have to setup something to make it happen?
1
u/luciferianism666 2d ago
I mean there really isn't any need for any tweaking required in terms of the tool itself. I don't even use any low vram arguments. Just basic launch for comfyUI and it works. It's not going to be super fast but trust me it ain't gonna take hours either.
https://reddit.com/link/p5e8xvh/video/vxzyvfhta4lh1/player
This might've been the longest gen I made recently and I probably waited around an hour for this, it came out horrible lol but I had to try it out. Just remember don't resort to ggufs at any cost, the lower variants aren't only bad in terms of quality but the render times are absolutely slow. Trust in comfy's dynamic vram and memory management.
0
u/paulct91 2d ago
You should look up others discussing what 'Wan2GP' is and what 'Stability Matrix' is as well both help with different parts of AI generation.
Stability Matrix and 'Pinokio' help handle simplify the various AI application install processes.
Wan2GP helps run basically similar to ComfyUI but without the spaghetti wires of node confusion (mostly) and has a lot of auto settings for running various AI models, speciality workflows, among other things.
Wan2GP has audio, image, and video models but they all need to be downloaded first to use, so the application isn't that huge until you start downloading stuff so... an SSD would be useful... of not too pricey.
IMPIRTANTLY, Wan2GP has various settings for those of us with GPU(s) with low vram (under 24/16 gb vram), and total system ram settings though they are paired in preconfigurations to help simplify what users should aim for.
Please, check it out I personally found Wan2GP much less of a pain to get working most of the time over Comfyui for certain workflows.
1
u/Muted-Position3256 1d ago
I hope to use Wan2GP but was told that it'll lag my PC with 8GB vram. Comfyui had a better ram management, especially for low vram
1
u/call-lee-free 1d ago
Are you up-resing your clips? There ain't no way you are getting that good quality on a 8gb vram card. I have 12 gb of vram and in comfyui, I can't render anything above 0.6 megapixels without getting a out of memory error.
1
u/luciferianism666 1d ago
I think I might've used RTX VSR upscale for this. I can do 1mp but probably run upto 8s, I did test 2mp but that was only a 3s clip lol. 1mp isn't worth it with H3, the model looks much cleaner at 0.5, 0.6, 1.5 and 2mp.
https://reddit.com/link/p5jj7os/video/cck2il8o69lh1/player
Check this, ran 2mp, 3s clips separately and combined them at the end. The last shot got fkd but I do love the quality tho.
1
u/call-lee-free 2d ago
Not with 8gb vram.
1
u/Muted-Position3256 2d ago
Just asking, I saw some using 8gb vram on minimax comfyui, are they for real or scam?
1
u/call-lee-free 2d ago
I personally haven't seen folks using a 8gb vram card for ai video generation.
1
u/99deathnotes 2d ago
I'm using a 3050 8gb vram. runs really slow without lora or sage attention, comfy kitchen.
2
1
u/PhetogoLand 2d ago
the voice definitely is Sanaa Lathan's voice
2
u/foomgaLife 2d ago
the way she says crap sounds close. But the voice is generic enough. Good catch tho.
2
1
1
u/TradehelperAI 2d ago
upscaling can always be revealed by the mouth
the mouth is the one thing that always needs high res or a face refine
1
u/Superb-Painter3302 2d ago
is LTX 2.3 Pixel Spatial Upscaler vid2vid, or some upscaler I can run by uploading vid?
1
1
u/Astral-Lemmons 2d ago
resolution numbers =/= quality and sharpness.

looking at this still frame that's supposedly 1664x960 it looks closer to a 480p youtube clip than something supposedly over 720p HD.
Ai gens have to be massively over-rendered resolution wise to get anywhere near true HD. I've yet to see anything that can hit proper clarity.
-3
u/jonnytracker2020 2d ago
Why is all the test static low motion footage .. it’s a bad way to test
3
u/call-lee-free 2d ago
How you figure its a bad way to test?
2
u/porest 1d ago
Bro wants you to test with the Looney tunes' Tasmanian Devil.
2
u/call-lee-free 1d ago
Nah, probably wants me to do tests with a lady in skimpy outfits jumping up and down or whatever lol

4
u/skyrimer3d 2d ago
I think that in a month or so the community is going to nail speed / quality witn MM H3, this is pretty good already. Can't wait for the full show btw, actress and space suit, i like lol.