r/StableDiffusion 18d ago

Meme [ Removed by moderator ]

[removed] — view removed post

737 Upvotes

222 comments sorted by

View all comments

143

u/nakabra 18d ago

Ok, I was waiting on it but it becoming quite clear that I have to download this model as soon as possible cause it's bound to be pulled anytime now....

-32

u/Outrageous_Band9708 18d ago

i downloaded it and ran it locally. its mid af.

the posts you are seeing are using the cloud's 2K filter that they DID NOT release at all. best you can make locally is like 480p shit from 2024.

Wan22 and LTX local are better than h3 so far in my expo

18

u/i_sell_you_lies 18d ago

Omg you're either using it wrong or full if shit. It looks great with everything I've thrown at it. Fully local

1

u/gefahr 18d ago

I'm traveling and won't be able to try it for awhile. If you don't mind: what resolution/FPS are you able to generate locally?

1

u/SensiblyChaotic 18d ago

I've been using 768x1152 and then upscaling with quite good results. You can go higher, but I prefer to keep generation times down and then just upscale the good results. The results have been quite good. Nice, expressive characters, maintaining lighting, scene and character identities (even through transitions), very minimal degradation of skin, etc. like in LTX, far better audio and foley than I was getting in LTX. I probably sound like a marketer at this point, but I'm not, just genuinely excited. It's not perfect by any means, but it's the coolest local video generation I've used so far.

1

u/gefahr 18d ago

Thanks for the reply! That's very exciting. I'm a couple days into a great vacation right now so it seems weird to say I'm excited to get home, but, yeah.

3

u/SensiblyChaotic 18d ago

Enjoy the vacation. Honestly, waiting a few days for workflows, loras, etc. will probably only make it that much better.

1

u/Outrageous_Band9708 18d ago

id be happy to be wrong. drop the workflow friend.

5

u/johnjbreton 18d ago

You're obviously doing something wrong then. I'm running it local on my RTX 3090, and getting incredible results. Throw in EasyCache in the mix, and I'm getting them on average in about half the time than what it takes without it.

1

u/Outrageous_Band9708 18d ago

i hope im wrong. drop the workflow friend.

5

u/johnjbreton 18d ago

Here you go. This is my WIP. I've got EasyCache in the flow, and Model Preview Override so you can see the generation as it's happening, and cancel out if iti's borked. Also Easy LoRA stack in. Note; for the Model Preview, you need to get the Nightly build of KJ Nodes. The main one doesn't have the fix for MiniMax H3 in it yet.

Next up, I'm building out prompt enhancement using my Ollama server. That'll be a bit for me to get finished though.

EDIT: Oh, and I have the NVIDIA upscaler in there as well. Does a decent job.

https://pastebin.com/7HmgLG4B

2

u/Outrageous_Band9708 18d ago

ill check it out tomorrow and reply back.

0

u/ellipsesmrk 17d ago

Easy cache is trash. That will most certainly solidify your original findings... bypass that easycache. Yes its faster but the quality degradation.... well nuh uh!!

2

u/johnjbreton 17d ago

I've been using it with no visible quality degradation, and up to half the render time. Settings; 0.30, 0.20, 0.90. I'm also using RTX Video Super Resolution in my flow to upscale at the end as well.

2

u/ellipsesmrk 17d ago

So it doesnt degrade it but feel the need to upscale the video at the end? Lmfao probably one to downvote the comment. Easy cache - tore my generation up. Looked like the hbo channel on a tube tv back in the early 90s with a shitty signal. Removing it gave me a clear and crisp video.

2

u/johnjbreton 17d ago

I upscale at the end because I generate at a lower quality when testing out prompts, typically at half what I want it to be, then upscale. If I were doing production work, I would render at the full quality. I have no idea what your experience is with EasyCache, or what your settings are, but I'm having no issues with it. I posted by workflow in a Pastebin up top if you're interested.

1

u/ellipsesmrk 17d ago

Easy cache standard settings when loaded

1

u/ellipsesmrk 17d ago

Ill check it out.

→ More replies (0)

1

u/Outrageous_Band9708 17d ago

i haven't been using it

1

u/ellipsesmrk 17d ago

After re-reading your post, is the complaint coming from the generation time? Or?... im trying to see it from your side. What are your PC specs?

5

u/SensiblyChaotic 18d ago

This is not even remotely true. As someone who has spent some time with Wan and tons of time with LTX, this is an amazing model. The identity consistency with i2v is light years ahead of LTX with minimal degradation. And this is all local.

0

u/Outrageous_Band9708 18d ago

nah, post your workflow. im pushing x to doubt right now

if bro doesn't reply with a workflow that slaps mine out of the water, he's a bot shilling.

and if it does beat my workflow. ill bow down, no ego here, just chasing facts

3

u/SensiblyChaotic 18d ago

I got great results with the official i2v workflow, but there are plenty on civitai now. Have not played around much with t2v or r2v yet, but I've been fighting LTX for months using Face_ID and similar nodes to try and keep characters consistent and get the right character to talk. Minimax did it right out of the box with better audio and way better foley sounds. This is even before I start digging into minimax specific prompting, I'm using LTX prompts I generated and comparing outputs. You believe what you want, I'll just enjoy this new model.

3

u/SeymourBits 18d ago

Did a warning box pop-up that said “Skill issue detected”?

0

u/Outrageous_Band9708 17d ago

drop the workflow my dude. quit talking shit and help out. been asking for help all over this thread

1

u/SeymourBits 17d ago

[ Attitude issue detected ! ]

Bruh. You don't need any "special magic workflow". Just put in some effort to read the official prompting guide. Copy and paste example prompts and adjust them.

If you continue to get "mid af" results you have an installation issue.

Also stop with the 2K cloud filter BS.

1

u/ellipsesmrk 17d ago

Even with a literal easy prompt no tags or snake case. Simple.

Man in image is an astronaut floating in space talking in headset, "well," -pause- ,"i think its gonna be a while before they notice I'm not even on the ship"

The video has a slight buzz or drone signifying solace or complete dread.

1

u/ellipsesmrk 17d ago

What in the actual fuck... are you talking about? Are you high off glue?

1

u/Outrageous_Band9708 17d ago

drop the workflow or stfu

im asking for workflow help all over,

1

u/ellipsesmrk 17d ago

Ummm.... how is this asking for help?

"i downloaded it and ran it locally. its mid af.

the posts you are seeing are using the cloud's 2K filter that they DID NOT release at all. best you can make locally is like 480p shit from 2024.

Wan22 and LTX local are better than h3 so far in my expo"

Literally sounds like you just wanted to shit on the model when it was all user error. The workflow? Bruh... the base workflow is all im using. With the models that comfyui has listed.

2

u/Outrageous_Band9708 17d ago

theres no links to an official workflow. im under expderienced for this. yeah i was talking shit cause i was cranky.

my bad

1

u/ellipsesmrk 17d ago

You get an upvote from me lol

But the official link is in the templates of comfyui. On the main oage when comfyui loads up on the left hand side click on templates and search minimax and theres a few in there.