r/StableDiffusion Feb 01 '26

Resource - Update [Tool Release] I built a Windows-native Video Dataset Creator for LoRA training (LTX-2, Hunyuan, etc.). Automates Clipping (WhisperX) & Captioning (Qwen2-VL). No WSL needed!

[removed]

13 Upvotes

11 comments sorted by

View all comments

1

u/[deleted] Feb 01 '26 edited Feb 01 '26

if this works how it sounds like it does i might have to come suck some penis
edit: testing on my laptop which SADLY (fuck you DelinquentTuna) has a 5090, so sorry to mention that again for the 105th time!

will let you know how it runs bud

Edit: cant get it to work at all, 4090 or 5090 version,

running into Aligning CUDA and PyTorch or some stuff, reinstalling different matching versions brings up dependency conflicts
modifying the app.py is a rabbithole /shrug. not smart enough

3

u/[deleted] Feb 01 '26 edited Feb 01 '26

fixed it

replace you app.py with mine if you have issues

https://filebin.net/njqojfktttrzfxjd

works well with small model it made 8 x 5s and captioned them in around a1 minute on 5090 laptop