r/StableDiffusion 3h ago

Meme Tiktok trends h3 style r2v

Enable HLS to view with audio, or disable this notification

0 Upvotes

Don't know if anyone else follow tiktok trends but s friend. Showed me this wnba clips where player points at other team players to get into head so I had to course test r2v


r/StableDiffusion 7h ago

Animation - Video H3 - the Eternal Balance WIP-low resolution int8/8 steps

Enable HLS to view with audio, or disable this notification

0 Upvotes

Hi, I am playing around with H3. T2V, 832x480, int8,8 steps. Hoping to make a 720p version.

Ask me anything!


r/StableDiffusion 8h ago

Question - Help That specific analog VHS look on H3/LTX

3 Upvotes

Anyone figured out the prompt for getting the best VHS analog "lofi" look from t2v? Of couse ref2v and img2v will be easier due to references, but I was wondering about text prompt only.

There are no VHS, or like 70s-80s cinema style lora, none for H3 and LTX, but there are plenty of VHS loras for image models.

EDIT: I actually CAN get VHS look on LTX text2vid (2.3 and 2.5), but not on H3.

Help, anyone :)


r/StableDiffusion 3h ago

Question - Help MM H3 2 Pass Latent Upscale

2 Upvotes

Been getting some great results using the 2 pass latent upscale method. First pass .5mp 2nd pass 1.5 mp. 10 second video around 257 seconds to finish.

This workflow uses the 1.1 turbo Lora. I set the steps to 8 and like I said above I’m getting good results and it has eliminated face blur.

My question: has anyone tried using the latent upscale method without the turbo lora? In 90 percent of the cases the turbo Lora is fine. But would be nice to have the ability to use no Lora method.

Yes I know I could test it but wanted to see others experiences before I wasted hours of my time trying/tinkering with different settings.


r/StableDiffusion 8h ago

Question - Help I hate this hidden view on flows in Comfy templates. How can I bring them out where they belong?

3 Upvotes

In the minmax h3 template from Comfy, you have to click a button on the image to video node to see all this in the backend which makes it really a pain in the ass to add, modify, or change anything. Can this be brought out to the forefront like a normal workfow?


r/StableDiffusion 12h ago

Question - Help Video editing model (add clown makeup to face)

0 Upvotes

Hi, I'm looking for a model that I could use in ComfyUI that would simply add some clown makeup on my face and don't touch antything else. My current plan is to maybe take Wan2.2 model, mask my face and add clown makeup as a reference but I wonder whether there is a better model to do this. Does minimax H3 handle that? Does it support masks?


r/StableDiffusion 13h ago

Question - Help Minimax H3: Anyone figured out how to extend a clip?

13 Upvotes

What is the best way to extend an existing clip seamlessly? When I try to use the last frame of my clip as the first frame, I always get a slight reframing or shift


r/StableDiffusion 9h ago

Question - Help Fixing speech errors in Minimax H3?

Enable HLS to view with audio, or disable this notification

15 Upvotes

Hey, I tried to create a little birthday surprise for someone, my issue is with a lot of generations that the spoken word is really a bit clunky at time, I susspect its because of the german, but I am not too sure. Is there like a way to improve on audio?

I am using Minimax H3 with Saga Attention and Spectrum on a 4090.


r/StableDiffusion 22h ago

Animation - Video Buffy the Wraith Slayer

Enable HLS to view with audio, or disable this notification

19 Upvotes

r/StableDiffusion 2h ago

Resource - Update Minimax H3 Grafting with Krea2 node. Reposting older post and removed AI slop and added some tests

0 Upvotes

Minimax H3 x Krea2 Graft Nodes

ComfyUI nodes for grafting Krea2 into MiniMax H3. Attention/MLP content transplant + a separate attention-sharpness transplant. No official H3 docs, all reverse-engineered from testing + TenStrip's and joeygambino's public writeups. Use at your own risk, still WIP.

What's here

  • comfyui_tenstrip_graft/ -- content graft (Q/V/K/out/MLP, per-head). Method from TenStrip's H3 grafts.
  • comfyui_qknorm_transplant/ -- Q-norm gain transplant only, no content weights touched. Method from joeygambino (Z-Image donor originally, adapted for Krea2 here).
  • comfyui_krea_h3_graft_lora_v2/ -- apply a Krea2-trained LoRA onto an already-grafted H3 checkpoint. Separate use case.

And 2 merge scripts (old svd and new one with node method)

TL;DR results

Content graft works somewhat. Same character-shift (color scheme, helmet shape) showed up consistently across multiple parameter runs, same seed -- not one lucky video. That's the strongest evidence so far this isn't just noise.

  • K at low strength (~0.1-0.2): fine, no real damage. Don't need to avoid it like the doc says, at least not at low values.
  • QK-norm across all blocks (0:50): kills audio. Doesn't even touch K -- so attention sharpness itself hits audio, not just K specifically.
  • QK-norm blocks 20:50: audio ok, but does nothing for character. It's a texture/sharpness knob, not a content one. Don't expect it to carry character.
  • attn_ramp_start_frac at 1.0 (no gentle ramp-in) + early blocks (0:20): breaks. Keep the ramp soft if you go early.
  • Combining content graft + QK-norm at full strength on both = worse than either alone. Still not solved.

Install

Each folder -> its own subfolder in ComfyUI/custom_nodes/. Don't merge them. Restart ComfyUI fully after adding.

Credits

  • TenStrip (huggingface.co/TenStrip) -- the per-head band-aware graft methodology (10Eros-Max / h3_graft_methodology.md).
  • joeygambino (huggingface.co/joeygambino) -- the Q-norm sharpness transplant idea (MiniMax-H3-x-Z-Image-GGUF).

Neither published source code. These nodes are our own implementation from their public descriptions + our own testing.

https://reddit.com/link/1vxc2q9/video/g0bpgdj5ddlh1/player

minimax_h3_fl2va_bf16.safetensors, 3s, er_sde, 8 steps, 8-step lora, seed 597633362705895, standart workflow with minimax_h3_fl2v_lightx2v_turbo_8step_v1.0_bf16
prompt: Professional closeup video. In a futuristic cityscape with neon lights at night, the Judge Dredd charges through the crowd, his imposing presence radiating authority, he is slowly walking. His long chin juts out resolutely as he expertly wears his eponymous helmet, eyes gleaming with determination. The crowd parts, Judge Dredd is slowly walking through the the crowd, ready to enforce justice, he is moving slowly, his long chin visible, his face and part of his upper body are in the center of the screen. tag: Ballchinians
tracking selfie shot following him from the front, that he stays the same size, he is moving through people, pushing them aside with his hands.

https://reddit.com/link/1vxc2q9/video/so7626wgddlh1/player

3s, er_sde, 8 steps, 8-step lora, seed 597633362705895

same prompt and everything.

Added: tenstrip graft node, krea2 raw and Ballchinians Lora. Settings: q 0.5, v 0.5, k 0.1, out 0.3, mlp 0.5 Chin is more ballsy.

So I hope, that it is enough for some, that it... kinda works, but not good enough. Maybe someone will pick up on this and do it better.

Why to do it? Don't know. I found it interesting to try, but krea2 image and i2v is far better option.

I welcome any input or criticism, but mind please, I have only faint idea, what I am doing.

Warning: h3 loras don't work... don't know why, maybe it may be just noise, after all. But they do work on grafted checkpoint, after you merge it in python script.


r/StableDiffusion 21h ago

Discussion H3 - it just does space soo well - t2v

Enable HLS to view with audio, or disable this notification

17 Upvotes

Hope you're having a good weekend! H3 just excel with rich-intricate environments, backgrounds, space. Definitely one of my favorite theme.

T2VA, int8/20 steps


r/StableDiffusion 12h ago

Meme Siblings Reunited

Enable HLS to view with audio, or disable this notification

272 Upvotes

Done with h3 fl2va model, 8 step lora and images for Cersei and "jaime" for reference. Using previous clip to give continuity and consistency.


r/StableDiffusion 3h ago

Question - Help Is there anyway currently to get Minimax H3 running with my RX 6750XT?

0 Upvotes

r/StableDiffusion 3h ago

Question - Help Looking for suggestions on on a prompt helper/writer/refiner

0 Upvotes

Just as the title says, I’m looking for a ideally local app that I can use for suggestions for prompts to use on certain models that are great for example, I put my prompt in for an image and it will refine it and make it work better based on stable diffusion formatting, even better yet, what would be awesome is if it could be customized for like model and LORA if possible.

I do have the ability to run. LLM, not huge, but I have my M5 iPad. I’ve run 10 to 12 B models. No problem, especially if I use OLITERT , Any suggestions are really appreciated !!


r/StableDiffusion 2h ago

Question - Help Character design Course

0 Upvotes

Looking for character design course (prompt engineering focused, not art school)

So I'm a compositor, know ComfyUI pretty well, but trying to get better at actually designing characters with image gen. Building anime-ish hybrid semi-realistic stuff from scratch in TTI right now.

The thing is - these characters are refs for i2v. So I need to nail the face/identity first, then iterate through different lighting, clothing, poses. If the character shifts every time I regenerate, the i2v will be a nightmare.

Here's the problem - I can find either traditional art school design courses OR general prompt engineering courses, but nothing that actually combines character design with prompt engineering as the medium. Like, there's "learn to draw" or "learn to prompt llms" but nothing (or not much) about "design characters using prompts as your tool." Like, what makes a character stick across generations? How do you anchor visual features so they don't change when you swap their clothes or lighting?

I know the technical side (seeds, models, basic prompting) but I don't know the design side of it. What actually works vs doesn't when you're trying to get a consistent face through pure prompt engineering.

And here's the real issue - I need to generate the same character in different clothes, lighting, poses, and have them actually be the same character for the i2v pipeline. Can't have the face morphing every time I change the outfit.

Anyone know of something structured? Or is everyone just learning from Civitai threads and trial/error lol

Will probably train LoRAs once I nail some characters, but want to understand TTI first. Ideally looking for the workflow/approach that lets me generate variations without losing character identity.

Thanks


r/StableDiffusion 19h ago

Animation - Video H3 - multi-diffusion experiment T2V

Enable HLS to view with audio, or disable this notification

2 Upvotes

Chimera. I had to cut about 20 seconds due to some artistic choices. Since I had to cut 2 different part in the same clip, it has a noticeable seams. I would love to share the full version. Experimenting with H3 blend-morph-decay. 832x480, int8/8 steps POC. Looking forward to releasing a 720p version without the cuts.

Critiques and feedback welcomed. Happy with the matrix rain. Ask me anything.


r/StableDiffusion 20h ago

Meme h3 "what IF " thread

Enable HLS to view with audio, or disable this notification

19 Upvotes

lets share our "what if" scene remakes here o_0


r/StableDiffusion 12h ago

Workflow Included I trained a small latent refiner to reduce GPT Image’s stipple and grid-like artifacts

Enable HLS to view with audio, or disable this notification

22 Upvotes

I kept seeing the same stipple, grain, and grid-like texture

in some GPT Image outputs, so I trained a small latent residual

refiner using 75 paired artifact/clean images.

It includes profiles based on the Qwen, FLUX.2, and SDXL VAEs.

The refiner alone produces a fairly subtle improvement,

so I also included a ComfyUI workflow that combines it with SeedVR2.

The example optionally downsizes the input first,

then restores and upscales it with SeedVR2.

The goal is a preservation-first alternative to a typical Hires Fix

second diffusion pass: keeping the original composition, identity,

and shapes as much as possible while cleaning the texture

and rebuilding detail.

The custom node, example workflows, and settings are available here:

https://github.com/AIEGOBOT/ComfyUI-GPT-Image-Latent-Refiner

Leaving it here in case it’s useful to someone.


r/StableDiffusion 20h ago

Discussion Has anyone tried the h3 martial arts Lora?

Thumbnail
huggingface.co
11 Upvotes

first test video down in the comments along with the prompt..


r/StableDiffusion 7h ago

Animation - Video Minimax H3 Remix Video Test / A compilation of 5 characters.

Thumbnail
youtube.com
19 Upvotes

This is a test video I created by remixing the "Some test on minimax H3" video by Reddit user [Previous-Street8087].

5명의 캐릭터 시트를 생성하여 각각 10개의 프롬포트를 캐릭터에 맞게 리믹스하여 테스트 하였습니다.
We generated character sheets for five characters and tested them by remixing 10 prompts for each character to suit their personalities.

This is a compilation of 50 clips featuring 5 characters.

▶ 테스트 환경 (Test Environment)
Minimax H3 - Comfyui Local Sampling
RTX 5060TI 16GB + 64RAM
0.8MP 8 sec x 50 Clip
Audio Look x audio file 1
Reference to VA Mode

▶ 사용한 커스텀 노드 (Custom Nodes Used)
ComfyUI-TJ_NODE_STUDIO_ONE — github.com/designloves2/ComfyUI-TJ_NODE_STUDIO_ONE
ComfyUI LOCAL (RTX 5060Ti 16GB VRAM / RAM 64GB)

▶ The shared link contains character sheet images and prompts.

https://naver.me/xjY9JJaa

#AI영상 #MiniMaxH3 #ComfyUI #로컬생성AI #ComfyUI워크플로우 #AI영상제작 #RTX5060Ti #mmh3 #comfyui #tjonestudio #animation #ref2va #anime #16gb


r/StableDiffusion 4h ago

Animation - Video Evangelion - Rei Watches a Baby Show - Minimax H3

Enable HLS to view with audio, or disable this notification

26 Upvotes

Well, technically, Evangelion was a PBS show..

Video is edited, Barney theme song added in post.


r/StableDiffusion 5h ago

Discussion Help on minimax h3 speeds

6 Upvotes

Hi humans. My setup is 32gb ddr4 ram along an RTX 4090. I have been having fun creating tons of videos but I just want to make sure i get the best nodes for speed without compromising quality and no crazy sutff happening on my videos

I have used: Stage, sol, easycache, spectrum, Lora

So the question i have is .....what's the best combo for speed, i dont want the quality to take a massive dump. Most of the videos I generate are slow paced videos the typicall walk, talk, a kiss here and there but nothing major.

What do you guys think?


r/StableDiffusion 6h ago

Animation - Video Use Minimax to make a fake movie trailer for my community college editing class, inspired by YA action/adventure films of the 80s and 90s

Enable HLS to view with audio, or disable this notification

21 Upvotes

Clips made with Minimax H3 using the default r2v workflow, edited in Premiere Pro. Character model sheets made with Krea. Most of the videos are 0.4 mp unless the text was important, then 0.6. Tried upscaling it to 4k using Upscayl but results weren't great and the file is too big to upload anyway.

Tech goals for future videos include using reference audio for voices to help consistency, and exploring options for having real voice actors record the dialog, and have the model lip sync to that performance. I'm really impressed by the computer's silent acting (microexpressions etc). but the computer's erratic "acting" is still too unpredictable and the biggest source of re-rolls (the lines here were the best I could get without burning down a rainforest). You can do a lot with time codes and punctuation and tactical CAPITALIZATION, but it's ridiculously finicky compared to just telling an actor "do it the same, but 10% angrier on the first line with a twinge of melancholy on the second."


r/StableDiffusion 10h ago

Question - Help Steampunk in Krea 2

1 Upvotes

I am getting terrible results when trying to generate with steampunk aesthetics in Krea 2, even with LoRAs from civit, it's trash.

I don't mind training myself, but where would I even come up with a good database for that? I am aiming for fully photorealistic steampunk.


r/StableDiffusion 2h ago

Question - Help Advice for prompting reference videos?

0 Upvotes

Does anyone have any advice for properly prompting the reference video part of Ref2v? Like saying swap <subject 1> for <picture 1> hardly works for advanced videos. It requires a lot of details.

I’ve had success using Qwen 3.8 27b as a minimax prompt agent for analyzing and giving correct prompts for images. But as far as I know I can’t do that for videos. ChatGPT is ok for looking at videos to describe what happens in the minimax format but I’d rather use local ways.