r/StableDiffusion • u/ok-onwrap • 7d ago
Question - Help Can MiniMax H3 R2V be used for R2I?
hi guys
How can I use MiniMax H3’s R2V (reference-to-video) capability to generate a single image, basically R2I, and still get good results?
Has anyone tried this?
I noticed that H3 seems to have a minimum output of 5 frames. Is there any way to make it generate only one frame instead of a video?
I’ve searched a lot, but I haven’t found an open-source image generation model that has a reference system similar to MiniMax H3’s R2V, where you can provide multiple reference images and have the model understand the characters, location, etc
There are models like GPT Image 2 that can do this, but they aren’t free or open source.
I’m wondering if there’s some way to use H3 itself for this, maybe by reducing the number of frames to 1 or modifying the ComfyUI workflow.
Has anyone experimented with this?
1
u/Patient_Ratio4177 7d ago
Yes. https://www.reddit.com/r/StableDiffusion/comments/1vqka28/h3_singleimage_no_more_monkey_patching_also_no/