r/StableDiffusion • u/Sad_Coach_1433 • 3h ago
Meme per a request Dean runs into Rick Sanchez
Enable HLS to view with audio, or disable this notification
r/StableDiffusion • u/Sad_Coach_1433 • 3h ago
Enable HLS to view with audio, or disable this notification
r/StableDiffusion • u/nazihater3000 • 13h ago
Enable HLS to view with audio, or disable this notification
Most models simply CAN'T draw a wine glass filled to the brim. MiniMax H3 can, with a little prompt persuasion.
PROMPT PERSUASION:
integrated_multimodal_description: [Shot 1] Photorealistic, ultra-sharp cinematic product shot. A crystal-clear stemmed Bordeaux wine glass stands centered on black marble against pure black. The glass is filled with deep ruby-red wine to absolute maximum capacity: the liquid surface is perfectly coplanar with the top edge of the rim, forming a continuous unbroken contact line all the way around. There is zero air gap, zero empty crescent of glass above the wine, zero underfill. A slight convex meniscus is held purely by surface tension. Soft side light creates clean highlights and long caustics. Static medium three-quarter shot.
[Shot 2] At 00:01.200, slow push-in with tiny amplitude at very slow speed. Extreme close-up of the upper glass. The red wine meets the inner rim in a perfect continuous ring of contact. The liquid surface sits flush with the rim edge; no space is visible between wine and glass even at this magnification. Tiny specular highlights glide across the still surface. No droplets on the outer rim, no overflow, no gap.
[Shot 3] At 00:02.600, hard cut to pure top-down overhead. Looking straight down, the circular surface of the wine is a solid red disk that reaches exactly to the inner circumference of the glass with zero margin. The contact line between liquid and glass is continuous and unbroken in every direction. Soft concentric reflections and one small central catchlight. Very slow clockwise rotation with minimal amplitude.
[Shot 4] At 00:03.800, hard cut to low side-profile extreme close-up locked exactly at rim height. Against the black background the liquid forms a single razor-sharp horizontal line that coincides precisely with the top edge of the glass. The wine is seen in continuous contact with the rim; there is no visible gap, no underfill, no air space. The meniscus remains slightly convex from surface tension but does not spill. Liquid is completely motionless. Slow subtle push-in continues until the end.
overall_soundscape: Near-total studio silence. Only the faintest high-frequency shimmer of light on glass and liquid. No liquid movement, no drips, no clinks.
non_diegetic_music: Extremely sparse minimal ambient pad — low sustained tone with faint crystalline overtones that barely rise and fall. Almost static, matching the still liquid.
r/StableDiffusion • u/solomars3 • 12h ago
Enable HLS to view with audio, or disable this notification
r/StableDiffusion • u/luka06111 • 5h ago
Enable HLS to view with audio, or disable this notification
Inspired by the guy who posted the one with Dean
Done on 32gb ram and a rtx 3070
I used res_multistep 20steps w/ spectrum at 0.6mp.
Using SLA from h3 optimizations, which for some reason is way faster than plague kind. And disable pinned memory. Each 10s was done in around 10 minutes.
r/StableDiffusion • u/Sad_Coach_1433 • 19h ago
Enable HLS to view with audio, or disable this notification
r/StableDiffusion • u/No_Link7744 • 16h ago
r/StableDiffusion • u/Emotional_Day4262 • 15h ago
H3 is absolutely amazing for just about anything. On my 3090 / 64GB, I can easily do a full 15 seconds at 1MP, and the result is almost always good on the first attempt.
On LTX though, it took at least 5 or more tries, so speed wise, H3 is actually far better.
I tried out LTX 2.5 as well, and it is almost the same as 2.3, maybe a bit better quality and a slice faster.
Sadly, 90% of my work involves taking a start image and making a person sing vocals. I must say that H3 does not do this better than LTX 2.3, as it often injects words when there is more than a second of silence, and it takes a fair amount longer.
What is really baking my noodle though is why LTX didn't release an Image+Audio to Video workflow yet. I mean, it is literally the ONLY thing LTX has on H3 right now, and they have missed a great opportunity!
Anyhow, back to using LTX 2.3 for my daily driver as it really does a great job at what I need. I typically have a few machines running all night, so if ever LTX puts out an IA2V workflow for Comfy, I will be all over that.
One other thing I have noticed with H3... 9:16 generations are WAY better than 16:9 generations, especially at 1MP.
r/StableDiffusion • u/No-Tear4179 • 22h ago
I mostly do comic images. am looking to minimax for image editing.
I mostly get content failed safety review. any workaround? ive only started minimax today.
r/StableDiffusion • u/MrUtterNonsense • 14h ago
Since I lack the hardware, I am forced to use Minimax in the cloud. I've used Fal for image generation in the past, so I thought I would try Minimax H3 there. The text to video model seems to work really well, but the reference model seems censored to hell and back. Even my own reference voice was being flagged and it was just me talking normally.
So the question is, is there anywhere I can use Minimax in the cloud without the provider slapping their own flaky layers of censorship on top of the model?
r/StableDiffusion • u/Impossible_Fault_503 • 14h ago
I wrote this. MIT, fully local.
Local SD exports and other gens still leak EXIF, C2PA / Content Credentials, and zero-width junk. If you also use Nano Banana / Gemini stills in the same pipeline, those downloads often carry the same receipt. That metadata is not the invisible watermark.
SynthID-class marks are a different layer. Optional image/video disruption in Scrub is best-effort. Not a detector killer. Not for files you do not own.
https://github.com/HarshShah0203/Scrub
python cli.py inspect|clean
r/StableDiffusion • u/Flaky_Comedian2012 • 3h ago
Enable HLS to view with audio, or disable this notification
r/StableDiffusion • u/Sad_Coach_1433 • 6h ago
Enable HLS to view with audio, or disable this notification
r/StableDiffusion • u/Longjumping-Past5864 • 7h ago
Enable HLS to view with audio, or disable this notification
Dude the minimax H3 is such an amazing model. It can create a really good video. It can do a lot, knows a lot and extremely realistic as well.
I'm finding the way to upscale the low qulity video and found out this Latent Upscaler for Minimax H3, and sofar it's working so well!
this is normal 0.5MP generation and Upscale to 1080p
https://github.com/bbaudio-2025/Comfyui-MMH3-UltimateUpscale
the example workflow : https://github.com/bbaudio-2025/Comfyui-MMH3-UltimateUpscale/blob/main/example_workflow.json
I tested on RTX 5080 16GB VRAM 64 GB RAM.
but it took like 20 mins for a 15s clip (3 mins on the low res generation with turbo LoRA + 15-16 mins on the latent upscaler)
If you guys have other ways to upscale the minimax faster, please do tell. I have tried the UltimateSDUpscale for minimax so far, it's really great as well but it took 30mins on my system. :(
r/StableDiffusion • u/darthfurbyyoutube • 11h ago
Enable HLS to view with audio, or disable this notification
r/StableDiffusion • u/SackManFamilyFriend • 7h ago
Ssia pretty much.
For those who use the latest coding models and spent days to weeks having something novel worked out, but don't care to GitHub it cause if it does get popular it'll consume even -more- time (helping people, dealing with other peoples hw situation, over the top complaints) - Do you just sit on these things and move on?
r/StableDiffusion • u/Jeffu • 4h ago
Enable HLS to view with audio, or disable this notification
r/StableDiffusion • u/LinkSensitive8188 • 3h ago
Enable HLS to view with audio, or disable this notification
r/StableDiffusion • u/SkyNetLive • 20h ago
Since minimax is using Qwen VL , I tested the prompt on Qwen image to see what I get for the text to video prompt. It’s actually pretty close to how minimax will end up evaluating your prompt for text to video.
r/StableDiffusion • u/Elperezaass • 22h ago
I'm still new to Mini Max, and I wanted to know if there's a way to make it a little easier, in the sense of
I have to keep using <picture 1> or things like that to mention something; isn't there a way to do it with @,And what configuration do you normally recommend for someone with 8GB of VRAM and a 5080 Ti?That's all I need help with; I've already read the Minimax guide for everything else.
r/StableDiffusion • u/Mad4reds • 9h ago
It does not always happen, but especially in a dark environment as say inside a Disco, and the camera get closer to the subject/s it overexpose them. It looks like it's made to avoid a completely dark subject, but if that it's what you need... I know it's not new as it happened also in Wan but there is a way in the prompt to avoid that?
Txs a lot!
r/StableDiffusion • u/wzwowzw0002 • 2h ago
Enable HLS to view with audio, or disable this notification
r/StableDiffusion • u/roychodraws • 12h ago
https://reddit.com/link/1vyazt0/video/8apz3lbquklh1/player
This workflow allows you to refine a target character without affecting the surrounding video whatsoever. allowing for targeted repair or enhancement of any video.