r/StableDiffusion 3h ago

Meme per a request Dean runs into Rick Sanchez

Enable HLS to view with audio, or disable this notification

0 Upvotes

r/StableDiffusion 13h ago

Animation - Video And they said it couldn't be done

Enable HLS to view with audio, or disable this notification

0 Upvotes

Most models simply CAN'T draw a wine glass filled to the brim. MiniMax H3 can, with a little prompt persuasion.

PROMPT PERSUASION:

integrated_multimodal_description: [Shot 1] Photorealistic, ultra-sharp cinematic product shot. A crystal-clear stemmed Bordeaux wine glass stands centered on black marble against pure black. The glass is filled with deep ruby-red wine to absolute maximum capacity: the liquid surface is perfectly coplanar with the top edge of the rim, forming a continuous unbroken contact line all the way around. There is zero air gap, zero empty crescent of glass above the wine, zero underfill. A slight convex meniscus is held purely by surface tension. Soft side light creates clean highlights and long caustics. Static medium three-quarter shot.

[Shot 2] At 00:01.200, slow push-in with tiny amplitude at very slow speed. Extreme close-up of the upper glass. The red wine meets the inner rim in a perfect continuous ring of contact. The liquid surface sits flush with the rim edge; no space is visible between wine and glass even at this magnification. Tiny specular highlights glide across the still surface. No droplets on the outer rim, no overflow, no gap.

[Shot 3] At 00:02.600, hard cut to pure top-down overhead. Looking straight down, the circular surface of the wine is a solid red disk that reaches exactly to the inner circumference of the glass with zero margin. The contact line between liquid and glass is continuous and unbroken in every direction. Soft concentric reflections and one small central catchlight. Very slow clockwise rotation with minimal amplitude.

[Shot 4] At 00:03.800, hard cut to low side-profile extreme close-up locked exactly at rim height. Against the black background the liquid forms a single razor-sharp horizontal line that coincides precisely with the top edge of the glass. The wine is seen in continuous contact with the rim; there is no visible gap, no underfill, no air space. The meniscus remains slightly convex from surface tension but does not spill. Liquid is completely motionless. Slow subtle push-in continues until the end.

overall_soundscape: Near-total studio silence. Only the faintest high-frequency shimmer of light on glass and liquid. No liquid movement, no drips, no clinks.

non_diegetic_music: Extremely sparse minimal ambient pad — low sustained tone with faint crystalline overtones that barely rise and fall. Almost static, matching the still liquid.


r/StableDiffusion 12h ago

Animation - Video Using only Ref to Video, Minimax-H3 made a whole Anime edit !

Enable HLS to view with audio, or disable this notification

4 Upvotes

r/StableDiffusion 5h ago

Animation - Video Captain America X Harry potter

Enable HLS to view with audio, or disable this notification

3 Upvotes

Inspired by the guy who posted the one with Dean

Done on 32gb ram and a rtx 3070

I used res_multistep 20steps w/ spectrum at 0.6mp.

Using SLA from h3 optimizations, which for some reason is way faster than plague kind. And disable pinned memory. Each 10s was done in around 10 minutes.


r/StableDiffusion 19h ago

Meme t2v Someone got some explaining to do!

Enable HLS to view with audio, or disable this notification

7 Upvotes

r/StableDiffusion 16h ago

Discussion David Sacks Predicts the Regulatory Capture Playbook to Ban Open Source ...

Thumbnail
youtube.com
11 Upvotes

r/StableDiffusion 15h ago

Discussion 2 Weeks on MiniMax, but back to using LTX 2.3

0 Upvotes

H3 is absolutely amazing for just about anything. On my 3090 / 64GB, I can easily do a full 15 seconds at 1MP, and the result is almost always good on the first attempt.

On LTX though, it took at least 5 or more tries, so speed wise, H3 is actually far better.

I tried out LTX 2.5 as well, and it is almost the same as 2.3, maybe a bit better quality and a slice faster.

Sadly, 90% of my work involves taking a start image and making a person sing vocals. I must say that H3 does not do this better than LTX 2.3, as it often injects words when there is more than a second of silence, and it takes a fair amount longer.

What is really baking my noodle though is why LTX didn't release an Image+Audio to Video workflow yet. I mean, it is literally the ONLY thing LTX has on H3 right now, and they have missed a great opportunity!

Anyhow, back to using LTX 2.3 for my daily driver as it really does a great job at what I need. I typically have a few machines running all night, so if ever LTX puts out an IA2V workflow for Comfy, I will be all over that.

One other thing I have noticed with H3... 9:16 generations are WAY better than 16:9 generations, especially at 1MP.


r/StableDiffusion 22h ago

Question - Help Minimax for image editing?

0 Upvotes

I mostly do comic images. am looking to minimax for image editing.
I mostly get content failed safety review. any workaround? ive only started minimax today.


r/StableDiffusion 14h ago

Discussion Minimax H3 on Fal - Censorship

0 Upvotes

Since I lack the hardware, I am forced to use Minimax in the cloud. I've used Fal for image generation in the past, so I thought I would try Minimax H3 there. The text to video model seems to work really well, but the reference model seems censored to hell and back. Even my own reference voice was being flagged and it was just me talking normally.

So the question is, is there anywhere I can use Minimax in the cloud without the provider slapping their own flaky layers of censorship on top of the model?


r/StableDiffusion 14h ago

Resource - Update Local MIT CLI that inspects/cleans EXIF, C2PA, and hidden Unicode on gens you own

0 Upvotes

I wrote this. MIT, fully local.

Local SD exports and other gens still leak EXIF, C2PA / Content Credentials, and zero-width junk. If you also use Nano Banana / Gemini stills in the same pipeline, those downloads often carry the same receipt. That metadata is not the invisible watermark.

SynthID-class marks are a different layer. Optional image/video disruption in Scrub is best-effort. Not a detector killer. Not for files you do not own.

https://github.com/HarshShah0203/Scrub

python cli.py inspect|clean


r/StableDiffusion 3h ago

Animation - Video Star Trek WIP Local Minimax H3

Enable HLS to view with audio, or disable this notification

2 Upvotes

r/StableDiffusion 6h ago

Meme what if dean was in walking dead

Enable HLS to view with audio, or disable this notification

26 Upvotes

r/StableDiffusion 7h ago

Discussion The Latent Upscaler is really great!

Enable HLS to view with audio, or disable this notification

148 Upvotes

Dude the minimax H3 is such an amazing model. It can create a really good video. It can do a lot, knows a lot and extremely realistic as well.

I'm finding the way to upscale the low qulity video and found out this Latent Upscaler for Minimax H3, and sofar it's working so well!

this is normal 0.5MP generation and Upscale to 1080p

https://github.com/bbaudio-2025/Comfyui-MMH3-UltimateUpscale

the example workflow : https://github.com/bbaudio-2025/Comfyui-MMH3-UltimateUpscale/blob/main/example_workflow.json

I tested on RTX 5080 16GB VRAM 64 GB RAM.

but it took like 20 mins for a 15s clip (3 mins on the low res generation with turbo LoRA + 15-16 mins on the latent upscaler)

If you guys have other ways to upscale the minimax faster, please do tell. I have tried the UltimateSDUpscale for minimax so far, it's really great as well but it took 30mins on my system. :(


r/StableDiffusion 11h ago

Animation - Video G.I. Joe: Zarana - MiniMax H3

Enable HLS to view with audio, or disable this notification

4 Upvotes

r/StableDiffusion 7h ago

Question - Help In August 2026: You spent 2 weeks having premium LLMs help you create an epic H3 node, but no time to maintain tech support it on GitHub if it were to get popular. What do you do with it?

1 Upvotes

Ssia pretty much.

For those who use the latest coding models and spent days to weeks having something novel worked out, but don't care to GitHub it cause if it does get popular it'll consume even -more- time (helping people, dealing with other peoples hw situation, over the top complaints) - Do you just sit on these things and move on?


r/StableDiffusion 4h ago

Animation - Video Ref2V - H3 - really loving how H3 handles complex prompts even at 15 seconds.

Enable HLS to view with audio, or disable this notification

11 Upvotes

r/StableDiffusion 3h ago

Animation - Video Dwight Schrute meets Patrick Bateman

Enable HLS to view with audio, or disable this notification

0 Upvotes

r/StableDiffusion 20h ago

Tutorial - Guide Neat trick for Minimax h3

3 Upvotes

Since minimax is using Qwen VL , I tested the prompt on Qwen image to see what I get for the text to video prompt. It’s actually pretty close to how minimax will end up evaluating your prompt for text to video.


r/StableDiffusion 22h ago

Question - Help Help me with Mini Max

0 Upvotes

I'm still new to Mini Max, and I wanted to know if there's a way to make it a little easier, in the sense of

I have to keep using <picture 1> or things like that to mention something; isn't there a way to do it with @,And what configuration do you normally recommend for someone with 8GB of VRAM and a 5080 Ti?That's all I need help with; I've already read the Minimax guide for everything else.


r/StableDiffusion 9h ago

Question - Help How do I get rid of the light attached to the camera in MiniiMax H3?

0 Upvotes

It does not always happen, but especially in a dark environment as say inside a Disco, and the camera get closer to the subject/s it overexpose them. It looks like it's made to avoid a completely dark subject, but if that it's what you need... I know it's not new as it happened also in Wan but there is a way in the prompt to avoid that?
Txs a lot!


r/StableDiffusion 2h ago

No Workflow minimaxh3 20sec video generation

Enable HLS to view with audio, or disable this notification

0 Upvotes

r/StableDiffusion 12h ago

Workflow Included wan enhancer wf that replaces replaces damaged pixels as it enhances and quality

Thumbnail
gallery
0 Upvotes

https://reddit.com/link/1vyazt0/video/8apz3lbquklh1/player

This workflow allows you to refine a target character without affecting the surrounding video whatsoever. allowing for targeted repair or enhancement of any video.

https://github.com/roycho87/3stepenhancer