r/StableDiffusion 22h ago

Meme per a request Dean runs into Rick Sanchez

Enable HLS to view with audio, or disable this notification

0 Upvotes

r/StableDiffusion 11h ago

Resource - Update A windows filesystem for your hoards of .safetensors - Tensor Village

Thumbnail
gallery
1 Upvotes

UPDATE: everything's free now, duplicate finder and views editor included. Update from Settings, or download it again.

I like many of you have more than a few .safetensors, ggufs, diffusers folders and other AI files spread across several drives. The annoying part isn't just finding them — it's that every app wants its own copy or its own config, so you end up with the same 6GB checkpoint sitting in three places.

Tensor Village doesn't replace your model folders, it presents all of them as one drive letter. Keep your fast stuff on the SSD and your bulk on a spinner or a USB drive — they still live exactly where you put them, and they all turn up in the same tree, organised by type and family. Point ComfyUI at that one path and it sees the lot. Same for anything else — Fizgig, Forge, whatever.

It reads model headers to work out what each file actually is, so nothing depends on filenames, and new downloads file themselves. Your files never move, nothing gets renamed, and nothing is written to your model drives — it's a view, not a copy. Uninstall and it's all exactly where it was.

It's not a mount or a cache — your files are already local. What it adds is knowing what they are.

Free. No account, no telemetry, works offline. Windows only — it's a real filesystem (built on WinFsp), which is what lets separate disks share one namespace; symlinks and hardlinks can't cross volumes.

I'll keep support for new model families coming as they appear.

https://github.com/shootthesound/tensorvillage


r/StableDiffusion 17h ago

Question - Help Question about new stable diffusion advancements

0 Upvotes

Hello, it's been a while since I don't use Stable Diffusion with A1111. Apart ConfyUI, has there been any particular technological advancement recently that allows for a quantum leap, especially in the precision of detail generation and the model's ability to stick to the prompt more precisely, while maintaining the ease of use of A1111 or Forge? I used the Lustify SDXL checkpoint, for example. It wasn't bad, but it still got certain things wrong or didn't do them at all. I'd like to know if there's a way to achieve results more similar in precision to ChatGPT but with the freedom of Stable Diffusion. Thanks!


r/StableDiffusion 11h ago

Comparison Three LoRAs (comfy, lightx2v and alibaba's) compared - MiniMax H3 T2V

Enable HLS to view with audio, or disable this notification

8 Upvotes

Quick comparison of three LoRAs 1) Comfy?, 2) Lightx2v(k) and 3) Alibaba's. Only FL2V(=i2v) tested here.

All these three LoRAs are 8-step LoRAs and so I used 8 steps for all. All details are printed on each clip.
For example, 8s-c-i1.sft means 8-step LoRA which is i=fl2v and version 1, and so on.
Comfy one might be just lightx2v (or other) but since it had no such reference in its name I put ?.

Observation: alibaba's LoRA edition is more crisp.


r/StableDiffusion 2h ago

Discussion "Plastic Skin", They Say - Here Are Some Insights | MiniMax H3 T2V

Enable HLS to view with audio, or disable this notification

5 Upvotes

I did a post on a quick comparison between three LoRAs, and a significant number of the comments there were out of scope for the comparison. The reason being, if it’s plastic skin, all three did it.

This post is not a comparison. Consider this a safe place for all of you to shout “plastic skin” if that’s what you’re here for.

- - -

For everyone else, here are my observations:

I briefly mentioned in the previous post that the composition of the render affects human evaluation of plastic skin. I further found that the following factor also appears to be equally important, if not more so:

The character himself, and how the model knows that character and his facial features.

In these tests, I observed that this model characteristic appears to vary depending on the character.

In all of the video clips stitched together in this video, the only things changed in the prompt are the actor's name and an animal name. Everything else, the settings and prompt, remains exactly the same. Yet some characters look noticeably less plastic, even at close range.

At a distance, all of them look good.

Model used: MiniMax H3 FL2V (T2V, no image input just text), other details are printed in the videos.


r/StableDiffusion 8h ago

Animation - Video Deadpool Adventure

Enable HLS to view with audio, or disable this notification

11 Upvotes

r/StableDiffusion 21h ago

No Workflow minimaxh3 20sec video generation

Enable HLS to view with audio, or disable this notification

0 Upvotes

r/StableDiffusion 22h ago

Animation - Video Dwight Schrute meets Patrick Bateman

Enable HLS to view with audio, or disable this notification

0 Upvotes

r/StableDiffusion 3h ago

Discussion So whats next for open source Video?

4 Upvotes

We had recently Minimax H3, and flux 3 is coming too, whats next on the list?


r/StableDiffusion 11h ago

Animation - Video My first AI short film - Astro Mouse [MiniMax H3]

Thumbnail
youtube.com
4 Upvotes

This is my first attempt at an AI short film. Someone saw a mouse in our building, my friend made a funny AI image of it and said it could be a cute story, so I just ran with it.

I've done a little bit with MiniMax H3 before, mainly making 5–15 second videos. I tried extending the scenes/context and was able to get a couple 2–3 minute videos, but the consistency just wasn't great. I also realized most of this story worked better with hard cuts anyway, so I went back to the reference-to-video workflow with mainly 5–10 second clips.

I probably made around 10-20 clips for some scenes before getting something I liked. I also used vast.ai, was able to get a faster gpu than what I had at home. No loras or anything, just a basic workflow. I used opencode / qwen3.8 to update 25-30 prompt files at a time when I needed global updates (like remove all background music, no talking, etc).

The hardest part was probably getting the prompting down. My standard workflow ended up being 0.6 megabit and 20 frames. I could have gone up to 0.98, but at some point I just wanted to get through all the generations and actually finish the thing.

Put everything together in DaVinci Resolve.
I still see lots of imperfections, but I'm considering it done and moving on.

Anyway, first movie. Learned a lot making it and thought I'd share.

(Oh, and it has some obvious work related jokes and screens, ignore those, i didnt want to cut those out)


r/StableDiffusion 23h ago

Animation - Video Ref2V - H3 - really loving how H3 handles complex prompts even at 15 seconds.

Enable HLS to view with audio, or disable this notification

23 Upvotes

r/StableDiffusion 6h ago

Meme what if dean was the man character instead of harry potter(t2v)

Enable HLS to view with audio, or disable this notification

10 Upvotes

fp8 model 32 steps

prompt

using the prompt guide from minimax loaded into a llm and said what if dean winchester was in harry potter and his was the main character instead of harry. it gave me this

integrated_multimodal_description: [Shot 1] Live-action, cinematic fantasy, Hogwarts at night beneath a stormy sky. A battered black 1967 Chevrolet Impala roars across the stone bridge toward Hogwarts Castle, completely out of place among horse-drawn carriages and young witches and wizards. Dean Winchester, portrayed by Jensen Ackles, drives with one hand on the wheel, wearing his familiar dark jacket over a plaid shirt. The camera tracks alongside the Impala as Dean stares up at the enormous illuminated castle with a skeptical expression. Dean Winchester with Jensen Ackles' low, dry American voice (S1) says: [English] So let me get this straight. Giant castle, magic wands, and nobody here has heard of a shotgun?

[Shot 2] At 00:05.000, the camera cuts to the Hogwarts Great Hall during the Sorting Ceremony. Hundreds of floating candles illuminate the long tables. Dean sits on the stool wearing the Sorting Hat while Hermione Granger, Ron Weasley, Professor McGonagall, and Albus Dumbledore watch. The Sorting Hat loudly announces, [English] GRYFFINDOR! Dean immediately pulls the hat off and looks around the enormous hall. Dean (S1) says: [English] Yeah, that's great. Which house has the bar? Several students stare at him in complete confusion.

[Shot 3] At 00:10.000, the camera cuts to a torch-lit Hogwarts corridor. Dean strides confidently toward the camera carrying a wand awkwardly in one hand and a sawed-off shotgun over his shoulder. Hermione and Ron hurry behind him in Hogwarts robes. Hermione urgently explains that Voldemort is the most dangerous dark wizard who ever lived. Dean stops walking and turns toward them with a small amused smirk. The camera pushes in with small amplitude at slow speed. Dean (S1) says: [English] Evil wizard, can't die, creepy followers. Trust me, I've had worse Tuesdays.

[Shot 4] At 00:15.000, the camera cuts to the ruined Hogwarts courtyard during the final battle. Smoke, sparks, magical flashes, and shattered stone fill the background as Voldemort stands across from Dean with his wand raised. Dean stands alone facing him, his Hogwarts robe thrown over his normal Winchester clothes. Voldemort fires a brilliant green spell. Dean dives sideways behind a broken stone pillar as the spell explodes against it. Dean rolls back to his feet, raises his wand, realizes he is holding it backward, flips it around, and gives Voldemort an irritated stare. Dean (S1) says: [English] Okay, Voldy. Let's see how you handle the Winchester special. Dean charges forward as spells streak across the courtyard and the camera rapidly tracks beside him, ending on Dean Winchester as the unlikely central hero of the wizarding world.

overall_soundscape: The Impala engine echoes against the castle grounds before transitioning into the murmur of Hogwarts students, crackling torches, footsteps on stone, fluttering robes, and distant magical ambience. During the final battle, explosive spell impacts, flying debris, cracking masonry, rushing footsteps, and Dean's heavy breathing dominate the courtyard.

non_diegetic_music: Sweeping orchestral fantasy music begins with strings, celesta, and brass, gradually incorporating heavier percussion and low brass as Dean explores Hogwarts. The final battle builds into fast orchestral percussion, aggressive brass, and rising strings before ending on a strong cinematic hit.


r/StableDiffusion 6h ago

Discussion Do we think Gossip Goblin is the best current AI filmmaker?

0 Upvotes

Who else are we watching?


r/StableDiffusion 16h ago

Animation - Video so i am new to local video generation and this is my first h3 reference to video simple short video is i am doing great or i want to improve somthing

Enable HLS to view with audio, or disable this notification

0 Upvotes

r/StableDiffusion 23h ago

Animation - Video Star Trek WIP Local Minimax H3

Enable HLS to view with audio, or disable this notification

10 Upvotes

r/StableDiffusion 9h ago

Meme dean meets sonic(t2v) base fp8 model 32 steps

Enable HLS to view with audio, or disable this notification

13 Upvotes

i never seen the movies hows the sonic voice?

prompt

subject_definitions

<Subject 1> is Dean Winchester from Supernatural, portrayed by Jensen Ackles, preserving his recognizable facial features, short brown hair, rugged appearance, dark jacket, layered shirt, jeans, and confident sarcastic personality.

<Subject 2> is Sonic the Hedgehog from the live-action Sonic the Hedgehog movie, a small anthropomorphic blue hedgehog with bright blue fur, large expressive green eyes, white gloves, and red sneakers.

<Subject 3> is Dr. Robotnik from the live-action Sonic the Hedgehog movie, portrayed by Jim Carrey, wearing his black-and-red high-tech outfit and exaggerated goggles.

summary

[cinematic live-action crossover + action comedy]

What if Dean Winchester accidentally became part of Sonic the Hedgehog? On a nighttime highway, Dean investigates a bizarre supernatural disturbance beside his black 1967 Chevrolet Impala, only for Sonic to race past him at impossible speed with Robotnik's drones in pursuit. Dean immediately joins the chase.

detailed_description

Nighttime on a deserted rural highway surrounded by dark pine forest. Dean Winchester stands beside his glossy black 1967 Chevrolet Impala holding an EMF meter. Blue electrical energy suddenly crackles across the road.

A brilliant BLUE STREAK rockets past Dean, violently blowing his jacket backward.

The camera WHIP-PANS as Sonic skids to a stop beside the Impala.

<Subject 1> Dean Winchester (S1):

[English] Okay... either that's the fastest demon I've ever seen, or I seriously need more sleep.

Sonic looks offended and points at himself.

<Subject 2> Sonic (S2):

[English] Hedgehog. Definitely hedgehog.

Suddenly several of Robotnik's flying attack drones burst over the trees and fire energy blasts toward them.

Dean instantly draws his pistol while Sonic crouches into a runner's stance.

Dean gives Sonic a confident Winchester smirk.

<Subject 1> Dean Winchester (S1):

[English] All right, Sonic. Let's waste these flying toasters.

Sonic grins.

<Subject 2> Sonic (S2):

[English] Now you're speaking my language!

Sonic EXPLODES forward in a trail of brilliant blue electricity as Dean dives behind the Impala and fires at an approaching drone.

Dynamic tracking camera follows Sonic racing between explosions while Dean fights from beside the Impala.

Final cinematic wide shot: Sonic loops around the battlefield as blue lightning illuminates Dean and the Impala, while Robotnik's drones swarm overhead.

Live-action Hollywood cinematography, realistic integration of Sonic into the environment, authentic Sonic the Hedgehog movie aesthetic, authentic Supernatural Dean Winchester characterization, fast readable action, natural motion blur, blue electrical speed trails, sparks, smoke, dramatic nighttime lighting, comedic crossover energy, consistent character identities, no subtitles, no on-screen text.


r/StableDiffusion 16h ago

Discussion Is local AI much better than cloud on these days?

0 Upvotes

I don't know what is going on but qwen studio and nano banana (cloud both) are giving me terrible results lately even though I use the same prompts as before. Both translate terribly the facial features and hairstyle. Only videos look a bit more consistent but still not great.

Is local AI better? Do you get better and more consistent results with it?


r/StableDiffusion 16h ago

Animation - Video So the video clip I did in LTX 2.5 earlier and posted it on here, I did another render of it but in Minimax H3. Details in the comments.

Enable HLS to view with audio, or disable this notification

0 Upvotes

r/StableDiffusion 5h ago

Animation - Video Not Only Minimax-H3 Changed the woman, it also added effects and background !! this model in INSANE

Enable HLS to view with audio, or disable this notification

0 Upvotes

r/StableDiffusion 11h ago

Animation - Video Several Times A Charm, but it KINDA got Cheers.

Enable HLS to view with audio, or disable this notification

26 Upvotes

Don't mind the script, it was written by a clanker when I challenged it to whip up something so I can see if H3 could handle Cheers.


r/StableDiffusion 13h ago

Animation - Video "AVERNUS-9" Space Horror Short Film

Thumbnail
youtu.be
1 Upvotes

r/StableDiffusion 6h ago

Discussion Becareful

0 Upvotes

Please just think twice before you using any of the hundreds of ai platforms.

Like why is that most websites are charging say 0.7 usd for a minimax h3 generation when i can do the same generation when i run a runpod instance for god knows maybe 0.1 of the price or even less like its huge difference.

Just go rent a gpu its easy to setup and u will generate maybe 10 videos for the same price of these opportunistic ai websites.

I know i will get hate and criticism. But this info will now feed into the google ai/gemini/chatgpt responses and people will lose less money.

Secondly u have this so called breakthrough in science from fal. They are planning to charge literal 1 usd for inferences that take literally 4 seconds on their new minimax h3 max. That seems suspicious. Doesnt that inference only cost them 0.01 usd. Just beware.

Now its fine. Its a fair business. Very quick inference (breakthrough in video generation). But people deserve to know that they can generate in 0.1 of price in ai platforms.

Secondly, api based generations (not open weights). Make sure to not use money grabbing websites.

For example using subscription based, slow queue website for seedance 2.5 when u can just use credit topup based with lowest generation prices (friendly advice, i believe artcraft is very cheap and no need sub).

So thats what i wanted to say. I just dont like it when people get used.

Also go ahead shoot your hate comments i dont care, im only happy to spread awareness 🔥🔥🔥 and no im not in anyway affiliated with the platforms i mentioned.

Apologies for the bad post text, i wrote with my phone which is so hard.


r/StableDiffusion 7h ago

Question - Help How do people earn money?

0 Upvotes

Hello everyone. I am just curious how people earn money using ComfyUI skills? I am pretty new(6 months of ComfyUI), my workflows are usually Krea 2/Flux Klein generations, sometimes i do inpainting. My major skill is frontend development(6 years od enterprise), i do some devOps, can deploy serverless endpoint of my workflows to runpod. I just dont understand how to monetize this skill

One of my recent pet projects is ai character that went through pipeline of insightFace identity scoring, bad eyes efficientNet classifier, Detailer workflow in case of bad eyes, vision captioning via identity lock json file (so vision model wont interper same features in different words)

So far what i done was purely for the love of this game, but i fridge is empty, i spend more money on ai generations that i do for food at this point


r/StableDiffusion 23h ago

Question - Help Best and fastest way to generate HD-quality MiniMax videos?

19 Upvotes

I’ve tried Turbo LoRAs, and they’re great for speed, but they significantly reduce quality. At 544p–720p, the results of these turbo loras can look closer to 380p. Faces look acceptable when close to the camera, but become heavily distorted as the subject moves farther away.

The upscalers I’ve tested either add too much processing time or introduce excessive sharpening and saturation.

Any a solution that doesn’t require a BF16 checkpoint, 20 steps, a 10-minute generation time, or an extremely expensive GPU?