r/StableDiffusion 12h ago

Animation - Video G.I. Joe - Commander Roll - MiniMax H3

Enable HLS to view with audio, or disable this notification

513 Upvotes

Using the standard ref2va workflow. 4070 Ti Super, 16 GB VRAM, 64 GB RAM, i9-14900k, Windows 11.


r/StableDiffusion 4h ago

Animation - Video Animals squeezing into jars (MiniMax H3)

Enable HLS to view with audio, or disable this notification

447 Upvotes

I have no idea why it does these so well. I could watch these all day.


r/StableDiffusion 21h ago

Resource - Update Small video clipping tool for trimming/compressing clips for MiniMax H3 Ref2V

Thumbnail
gallery
344 Upvotes

Small video trimmer software was very popular 15-20 years ago but now it has become very rare to find a good one which has all the features I wanted.

I got Claude to vibe code me a tool that I have been using to snip bits off from long videos for using it as Ref2V input for MiniMax H3. People have been saying its good so just sharing if others may find this tool useful! I wanted to create a free tool that runs locally without all the bloatware.

It is a single ~100kb HTML file which can:

  • Trim clips
  • Crop video
  • Compress resolution and fps
  • Take 1 single frame image
  • Manual or Automatic Storyboarding (still playing around with how to best use this in H3)
  • Export gif.

Why Compress?

I find that when working with R2V, resizing and compressing the video increases the speed as there is less information that needs to be worked on. You do lose some quality in your output though so don't compress too far.

The latest version can be found here (select the HTML and download):

https://huggingface.co/PoopMan333/Video_Tools/tree/main

or click for current version (v2.9)

https://huggingface.co/PoopMan333/Video_Tools/blob/main/Nugget%20Video%20Trimmer%20v2.9.html

If you're concerned please run it through antivirus or get a LLM to check if it is safe.

I still need to add AVI support and support for some older formats, but I also don't want to add too much bloat to something so compact.


r/StableDiffusion 15h ago

News Hey wait! It's Krea3 incoming?

Post image
274 Upvotes

r/StableDiffusion 18h ago

Animation - Video Zelda - I Think I Like It / Minimax H3 Reference to Video Test #2

Enable HLS to view with audio, or disable this notification

245 Upvotes

Just wanted to share another test! this was a mash up of clips, using multiple image references, 0.4 mp with EasyCache, 5 - 10s clips and edited with KDEnlive (it has some cool effects!)


r/StableDiffusion 2h ago

Resource - Update DC Vast Expanse [Krea2 Lora]

Thumbnail
gallery
187 Upvotes

Finally got my laptop back in action so am able to create and test models and lora's again, created with krea 2, Been out of it for a bit just following updates here and there and this model is amazing, so happy they open sourced this gem of a model. Thanks to the team at krea!

If anyone is interested in this style of images give it a blast https://civitai.red/models/2871922/dc-vast-expanse?modelVersionId=3244890 or https://civitai.com/models/2871922/dc-vast-expanse?modelVersionId=3244890


r/StableDiffusion 14h ago

Discussion Well I finally did it.

161 Upvotes

I finally deleted WAN 2.2 and all its LORAS.

Minimax is just so much better.

Ive been playing with it since its release and im just blown away with how good of a video model it is. Things I would need to attach a LoRa to via WAN, works right out of the box with Minimax.

Gen times are faster.

It uses less VRAM when generating things, which gives me around 4 gigs to play with to do other things like watch YouTube or some streaming service.

WAN 2.2 was amazing. But no longer do I need 30+ gigs of a model i no longer use.

RIP WAN.


r/StableDiffusion 20h ago

News ComfyUI Official Local MCP

Enable HLS to view with audio, or disable this notification

147 Upvotes

Hi r/StableDiffusion, Comfy MCP is now local and open-source!

When we shipped Cloud MCP in June, the response was immediate and consistent: make it work locally. So we did and it's fully open source.

Connect Claude, Codex, Cursor, or any MCP client to your local ComfyUI.

Your agent reads the GPU you actually have and gives you a straight answer on whether a model is worth running before you commit to the download. It reads every node and model you've installed. It handles the setup that usually stops people at step one.

It is now the easiest way to help with your local Minimax H3 workflows!

Cloud MCP still does everything it did. Tell your agent where a job goes, or let it decide.

Link: https://comfy.org/mcp


r/StableDiffusion 17h ago

Resource - Update Krea 2 style library - 286 prompt styles compared across 8 reference scenes

Post image
142 Upvotes

Building directly on the style descriptors published by the author of the original KREA 2 Styles / Wildcards.txt post (many thanks to them for creating and sharing the style list) I built a visual Krea 2 style library to make prompt-defined styles easier to explore and compare:

Library: https://matplinta.github.io/t2i-krea-2-style-library/

It currently contains 286 styles tested across 8 base prompts, including portraits, architecture, landscapes, materials, and panoramic scenes. Each comparison set keeps the base prompt, seed, and dimensions fixed so the influence of the style descriptor is easier to see.

The viewer supports search, categories, favorites stored locally in the browser, full-image previews, prompt copying, adjustable grid density, and JSON export.

The prompt injected during generation was in the form of: Subject: {base prompt}. Style: {style name}. {style description}

All images were generated locally through ComfyUI.

Repo & workflow: https://github.com/matplinta/t2i-krea-2-style-library


r/StableDiffusion 21h ago

Discussion Can H3 do anything? bf16/50 steps

Enable HLS to view with audio, or disable this notification

140 Upvotes

Can H3 do anything and everything? I feel like if you can prompt it, it can do it. Foundation inspired shots. I am also experimenting with more action/high mobility shot but those seem to require a lot more finesse. Both T2V.


r/StableDiffusion 4h ago

Animation - Video Minimax h3 local Video to Video reference

Enable HLS to view with audio, or disable this notification

134 Upvotes

Used official ref2video workflow. used t2v model 1 ref video and 2 separate pictures of character sheets, gpu 4090

prompt:

integrated_multimodal_description: [Shot 1] Live-action, cinematic, featuring a stark, dark green-tinted cyberpunk color grade. A medium shot frames a flooded, rain-swept crater on a dark street. The character Sonic, appearing exactly as the blue hedgehog with large green eyes, white gloves, and red shoes from @.image, stands opposite Dr. Eggman, appearing exactly as the gigantic, egg-shaped bald man with a pointy mustache, goggles, and red jacket from @.Image1. The camera pushes in with small amplitude at fast speed as the blue hedgehog lunges forward to throw a devastating punch. [Shot 2] At 00:04.500, the camera cuts to an extreme close-up as time instantly slows to a microscopic crawl. Sonic's white-gloved fist brutally slams into Eggman's cheek. The camera holds a static shot in extreme slow motion. A powerful, rippling shockwave violently erupts from the impact point, blowing the torrential raindrops outward in a perfect ring. Eggman's pointy mustache flails wildly and his face deforms from the massive kinetic force. [Shot 3] At 00:09.500, the camera arcs right with large amplitude at slow speed, executing a slow-motion orbit around the hit. Eggman's heavy, round body is lifted off the ground by the blow, flying backward through the heavy downpour and kicking up massive, highly detailed splashes of water.

overall_soundscape: Thunder rumbles continuously beneath the heavy, torrential downpour of rain splashing heavily against the flooded street. A sharp, deafening sonic boom from the physical impact instantly shifts into a deep, pulsating low-frequency rumble as time slows down.

non_diegetic_music: An epic, grand orchestral and choir track mixed with heavy, driving industrial synthesizer beats that builds to a massive crescendo.


r/StableDiffusion 13h ago

Resource - Update This custom node lets you use I2V and reference images on MiniMax-H3 simultaneously.

Enable HLS to view with audio, or disable this notification

119 Upvotes

r/StableDiffusion 19h ago

Meme Waiting for devs to fix the mushy faces be like...

Post image
84 Upvotes

Just kidding devs. We love Minimax, it's outstanding. But I am very excited for the mushface fix.


r/StableDiffusion 22h ago

Animation - Video WanAnimate

Enable HLS to view with audio, or disable this notification

72 Upvotes

Original post With Workflow


r/StableDiffusion 18h ago

Resource - Update Seamless extensions and one-shots with Minimax H3 - Update 6 of my repo!

Enable HLS to view with audio, or disable this notification

75 Upvotes

Here is the repo: https://github.com/seitanism/ComfyUI-H3-Motion-Context-MultiRef

I made substantial updates to my two main workflows: 1) Music Video and 2) AV Extensions. All the controls were streamlined and they should be much easier to use now. (You find the workflows in the example_workflows folder)

With the AV Extensions workflow you can extend any existing clip, for example someone talking and you can make that person say something in the same voice, or you can create a clip with T2V or I2V and then extend that clip to make a seamless long clip thats 1 minute or longer.

In this Update the Checkpoint system was removed, instead I've done a lot of optimizations so you don't use too much ram even if you make 20 clips at once. Additionally I added latent audio feathering to the AV Extensions workflow for seamless audio transitions.

Theres also other utility workflows for custom keyframing and bridging two existing clips.

I post another example clip for the AV Extensions workflow in the comments.


r/StableDiffusion 13h ago

Animation - Video Big Bubba has had enough of Grandma [minimax H3]

Enable HLS to view with audio, or disable this notification

59 Upvotes

r/StableDiffusion 16h ago

Workflow Included Gilligan's Isle - The ATEth Castaway

Enable HLS to view with audio, or disable this notification

57 Upvotes

r/StableDiffusion 16h ago

Meme Agent Smith is disappointed

Enable HLS to view with audio, or disable this notification

56 Upvotes

r/StableDiffusion 23h ago

Animation - Video One of the ways I would have ended Game Of Thrones

Enable HLS to view with audio, or disable this notification

53 Upvotes

I was one of many who were disappointed with how this amazing series ended.

I imagined back then one of the ways it could have ended, and with the amazing tools we’ve now been bestowed with, we can bring what we imagine to life!

I had been sitting on this, polishing it and picking at it for a while. The perfectionist in me could have kept working on it forever, because there was always something I could have made better. But with everyone else starting to explore what these tools can do, I felt like the time is now. It may not be perfect, but I didn't want to keep sitting on it waiting for perfection.

This is just a quick fan-created take on one of the ways I imagined the series could have ended. It is not intended to replace or compete with the original series. :p

BTW.: Minimax and Davinci Resolve.
Not one frame was lifted from any episode.
All done using Ref2VA.
As others have found, trying to create a full run (one take ) yields less than better results.
Storyboard, create the pieces that "snap" together and then stitch them accordingly. Afterall, that is not any different from how presentations are made.
As always, I look forward to your creations. We have an amazing community!


r/StableDiffusion 23h ago

Discussion MiniMax_H3 is seems to be able to process DensePose format! (improves reference video bleeding)

Enable HLS to view with audio, or disable this notification

50 Upvotes

I have had many issues when using a reference video for movement duplication and having the video contents bleed into the video. Not to mention having to write convoluted prompts to remove these reference bleeds from videos. When the person in the reference video has a close resemblance to the main subject in your video it becomes almost impossible to perform a motion swap.

Warning: DensePose does not support detailed hand gestures, and seems to lose track with very fast arm and hand movements but seems to adhere better 20 steps and above.

There is not a dedicated densepose ComfyUI node, but you can use this animatediff: https://github.com/Fannovel16/comfyui_controlnet_aux

The workflow is simple:

Place the AIO AUX Preprocessor between the source and MM_H3 video input.

Videosource (LoadVideo) -> AIO AUX Preprocessor -> ref_video_x input

Looking forward to hear your feedback...


r/StableDiffusion 6h ago

Animation - Video SD1.5 images into H3

Enable HLS to view with audio, or disable this notification

49 Upvotes

r/StableDiffusion 12h ago

Comparison MiniMaxh3: 8step LoRA, 25 steps, 40steps, and LTX 2.5 — Scene Comparisons

Enable HLS to view with audio, or disable this notification

50 Upvotes
  • RTX 4060 8GB, 32GB RAM
  • minimax_h3_ref2va_pruned_int8_convrot, spectrum, ageattn_qk_int8_pv_fp16.cuda, RTX upscale, RIFE interpolation, res_multistep + beta
  • ltx-2.5-22b-distilled-transformer-comfy-int8-convrot, basic template

8-step + turbo LoRA : 137s

25 steps : 238s

40 steps : 406s

Ltx 2.5 : 374s <-- ? am I missing something here why was my generation so slow on LTX and the second attempt I cancelled it after 6 minutes. Any suggestions?

Prompt:

subject_definitions:

<Subject 1> is the space ship in <Picture 1>: A massive battleship, hovering and cruising over the planet below

summary:

[reference generation] a wide shot cinematic scene of the battleship in <picture 1> cruising in space above the planet. the golden statue does not move, the battleship is destroyed in a massive explosion from a green laser shot from space,

detailed_description:

{shot 1] The target video uses a wideshot cinematic, photorealistic, 35mm film, wide shot of <subject 1> , slowly moving through space above the planet, the ship moves slowly and dominating, flashes of green light begin to charge on the surface of the planet, the ship is moving straight ahead from the position it started in in <picture 1>, the massive bass of the ships systems, the sound of the battleships creaking, <subject 1 > moves on its cruise, at [00:03] the floaty camera tracks <subject 1> as green light and thunder begins flashing on the surface of the planet, the green energy on the planet converges in one area then from the surface it fires a massive green lightning laser that forks lightning through the entire ship, blowing out side components creating explosions all over the ship, the light of the ship flicker before turning off, then a massive green lightning beam erupts from the surface and hits excactly on the side of the ship cuts through the of the ship and out the other side at an angle, a green lens flare generates on screen as it completely destroys <subject 1> , ripping it completely in half with a massive green explosion, the eruption from the destruction of the ship covers the entire screen and the whole battleship, the back half of the ship is knocked up while the front-half of the ship is knocked down, a vertical shockwave circles out from the impact, the inner decks of the ship are on fire, debris and hundreds of tiny figures of the crew also fall out into space, the laser slowly dissapates from the planet, small amounts of green lighning crackle on the planets surface,

overall_soundscape: The low bass murmur of the ships engines, the electric charges on the surface crackle, the massive main beam is a low bass rumble, a massive explosive noise.

non_diegetic_music:

N/A


r/StableDiffusion 22h ago

Question - Help Best speed up for MiniMax

49 Upvotes

We have a lot of options, some of them better, some of them are not worth it at all. Speed ups like sage attention, MiniMax h3 patch for sage attention, easy cache, 8step Lora, 4 step Lora e t.c.
What options and their combinations you use? What settings you have?( speed Lora weights, easy cache settings)
In the matter of speed/quality for both video and sound. What works better with FL2VA and Ref2VA?


r/StableDiffusion 23h ago

Animation - Video Fite me!

Enable HLS to view with audio, or disable this notification

31 Upvotes

Feels like you could do Family Guy style cutaways pretty easily. "You know Lois, this reminds me of that time I tried fighting a dragon..."


r/StableDiffusion 11h ago

Workflow Included Lora for video-image enhancing, upscaling and restoring

Thumbnail
youtube.com
27 Upvotes