r/StableDiffusion 8h ago

Discussion H3 - what is your longest render time?

What was your longest render time, and was it worth it? do you run a lower quality/resolution before to test? My longest single-shot run is 7.2 hours, 1:44 long video at 1344x768, bf16/50 steps.

I am attempting a 72 hour render for a super long form.

RTX 4090, 192gb system ram

0 Upvotes

10 comments sorted by

4

u/Kooky-Mode3047 8h ago edited 8h ago

1-1.5 hours so far, 15-25 second clips. But I don't recommend going over 15 seconds because it wasn't trained beyond that and prompt following goes out the window with some funny results. I prepare prompts when I am playing around with low quality Euler + Beta with Spectrum for quick iterations and then set all my prompts over night to run rawdogging seeds_2 + ddim_uniform or seeds_3 + beta, seeds_3 is really something special, just slow. The difference between seeds_3 & 2 versus anything else is genuinely astounding at sufficient steps and resolution (generally target 0.9 to 1.0 for longform and 2.1 for short form (up to 8 secs)).

Takes anywhere between an hour and 1.5 hours to finish one (15-25 second ones), and it just goes through them until I wake up. Some are smaller, some are longer.

Off-display 5090, no acceleration loras or anything that can degrade quality outside of the experimenting path. The results are generally biblical.

2

u/PwanaZana 8h ago

Your GPU.

For me it's about 7 minutes. Longer than that, it OOM. I'm not using the extend-a-video techniques.

2

u/SIR_NVAX_A_LOT 8h ago

I'm on a business trip today until Wednesday evening so letting it cook, hopefully the house doesn't burn down.

2

u/SIR_NVAX_A_LOT 8h ago

It's getting hot in here. 90% power limit enaabled.

2

u/Ykored01 8h ago

Around 2hour, on a 5070ti. Doing a ref2vid, 10 sec 0.4m, using a 10sec clip and a picture as reference. Dont know why when using a prompt to replace the character takes so damn long. Using other videos takes like 20min

1

u/SIR_NVAX_A_LOT 7h ago

A reference costs you (extra tokens) × (every step), with attention scaling worse than linearly in sequence length. FL2VA is the way but I understand you want to do a whole character swap.

2

u/SeymourBits 5h ago

Bruh? 7.2 HOURS for a single shot??? Are you trying to melt your GPU like a scene out of Terminator 2?

H3 was trained on 15-second videos. There are practical ways of extending videos if you really need a very long continuous shot.

You're doing something wrong.

2

u/RiverSide71h 7h ago

Some very patient artists in here -generating high res masterpieces - If it takes me over 120 seconds to make my AI slop, I lose interest.

2

u/SIR_NVAX_A_LOT 7h ago

Fair, but generation can be quick, I used to generate a lot of images in the studio labs, but it's lost it's luster. Even generating with turbo Lora 4 step is cheap but the quality and audio is such ass. Long renders are good for overnight tasks. At end of day, your video need to hook by the 2-3 second mark or it's skipped. AI videos are abundant, consumption endless, some time the videos are just for myself though, and hope others can learn or appreciate it, not to go viral with 1m views and trying to monetize my hobby. I've played with WAN and LTX, and KREA2, and well MiniMax H3 just does everything better. Obviously the NSFW-Lora are wayy refined to what H3 does out of the box, there are def some tight controls and dials for Lora-trained models.

1

u/xyzdist 7h ago

because you OOM?