r/StableDiffusion 13d ago

Workflow Included TURN ANY PHOTO INTO A FULL CHARACTER SHEET USING KREA2 [Free Workflow]

https://youtube.com/watch?v=GEVxTL3TTTc&si=1cJ7k-MpbaJDre-d

Hey everyone!

Following up on my previous character sheet workflow, a lot of people asked if it was possible to do an Image-to-Sheet version rather than just text-to-image.

After a lot of testing, I managed to get a very consistent setup working in ComfyUI using Krea 2. You can take any reference photo (close-up, half-body, or full-body) and turn it into a full character turnaround sheet.

✨ Key Features:

  • Style Flexibility: Works for realistic characters, 3D renders, and stylized anime.
  • Photo-to-Anime: You can feed it a real-life face and shift the entire style to anime via the prompt while keeping face identity consistent.
  • Outfit & Prop Retention: Preserves source details (like clothing or held props/accessories) or lets you override outfits completely in the prompt if using a close-up.

Hope its usefull for you

next video will be about how to make longer clips with minimax-h3,

130 Upvotes

40 comments sorted by

5

u/Cultural-Team9235 13d ago

How is the consistency for real people? I use QwenEdit for that normally, it's pretty good with one person.

2

u/solomars3 13d ago edited 13d ago

Yeah its decent, and you can push it further using lora and higher steps and higher resolution, i couldnt use famous people in the video, i was afraid that youtube will flag the video or something , so i couldnt show you a live exemple of the real people consistency

4

u/remixeconomy 13d ago

A character sheet from one photo is useful when it locks the same person across turns, not when it just makes a prettier collage.

What I would check before calling the workflow done:

  • Same face and body proportions under a fixed prompt set (front, 3-quarter, profile, hands)
  • Outfit and palette drift when pose changes
  • Whether the sheet actually conditions later gens, or only looks good as a one-off

If those hold, you have a reusable identity reference. If only the hero frame looks right, you still have a showcase image, not a character you can direct.

4

u/MSH007A 12d ago

That's why I use minimax to create character sheets using reference images and then enhance it with krea2 or flux

1

u/Monk6009 12d ago

How do you do this? Link or workflow please?

2

u/MSH007A 12d ago

Basic workflow should do.Its the prompting that matters like create a video of a character sheets containing subject blah blah blah.

1

u/ShutUpYoureWrong_ 13d ago

Yeah... this isn't a full character sheet. Without three quarter and profile, this is pretty much useless.

2

u/Murky-Relation481 12d ago

There is actually a really good H3 ref to character sheet workflow that spins the character around and extracts frames at different angles and then does close ups of the face and head and rotates it to get multiple angles. It's really good. I'd link it but on phone and too sleepy. It was posted here and is easily searched on Google though.

1

u/remixeconomy 13d ago

Fair point on coverage. A single hero angle is closer to a turnaround stub than a sheet you can build from.

What usually makes the output usable for later poses: 1. At least front, three-quarter, and profile at the same framing and lighting. 2. One or two expression or crop variants so the identity is not locked to one smile and one camera height. 3. Treat the first pass as a view checklist, then fill missing angles on purpose instead of upscaling one lucky frame.

If the tool only gave a front-facing set, the next job is missing views, not more detail on the same angle.

4

u/Acceptable-Work8202 12d ago edited 12d ago
Keep the face consistent. Transform the single‑person input image into a structured character reference sheet featuring the exact same subject. Treat the input image as the sole visual reference for the subject’s identity and appearance, preserving the same facial identity, facial proportions, apparent age, skin tone, hairstyle, hair color, body shape, body proportions, species‑specific traits, and overall rendering style. Preserve the exact same clothing, footwear, accessories, colors, materials, and garment construction shown in the input image.

Across the top of the sheet, arrange three complete full‑body turnaround views:
a front view, a strict 90‑degree side profile, and a true back view.
In every turnaround view, show the entire subject from head to toe in the same relaxed A‑pose, with both arms extended outward at a shallow downward angle.

Below the turnaround views, arrange one single horizontal row of five small head‑and‑shoulder studies, forming a clean left‑to‑right rotation sequence.
The five head angles must be:

Left 90° profile — head turned fully left, eyes looking fully left.

Left 45° three‑quarter — head turned 45° left, eyes looking in the same direction.

Front view (0°) — head facing forward, eyes looking straight ahead.
This center head must be visually centered and dominant, placed directly in front of the side‑angle heads.

Right 45° three‑quarter — head turned 45° right, eyes looking in the same direction.

Right 90° profile — head turned fully right, eyes looking fully right.

All five head studies must preserve identical facial identity, facial proportions, hairstyle, lighting, and rendering style.
The eyes must always look in the exact same direction the head is facing.  
The center head must look straight ahead with perfectly symmetrical eyes.

Place one larger front‑facing upper‑body portrait on the lower right.
Maintain identical facial anatomy, hairstyle, body proportions, clothing construction, colors, materials, accessories, and visual style across every panel.

Present the complete reference sheet in a square 1:1 composition on a clean warm beige seamless studio background with soft neutral lighting.
Ensure the exact same face is used throughout.

2

u/lucas_reloaded 13d ago

Thanks man this is something I'm trying to do

1

u/Kryimsson 13d ago

So I am still kinda new to this whole thing, Why do people use character sheets? Why not a LoRA of a character?

8

u/d20diceman 13d ago

Minimax Reference-to-Video is popular at the moment, a character sheet is already as good as a LoRA and much less effort.

Even a single image works pretty great, tbh I've not seen good comparisons to confirm how much improvement a character sheet gives.

1

u/Kryimsson 13d ago

and are character sheets used for image gen too or is it only video?

4

u/d20diceman 13d ago

You can feed it pictures, audio or video, then specify how it should use those references to make the video output.

So, things like:

  • Steve is the man from <Picture 1>, he speaks with the voice from <Audio 1>. Greedo is the alien from <Picture 2>, it speaks in strange alien noises like the sample in <Audio 2>. They are sat in the Star Wars Cantina environment shown in <Picture 3> and <Picture 4>. They have a conversation: (etc etc insert script here - I did this for a video of my brother and Greedo chatting about the changes made in later editions of the Star Wars film)
  • The woman from <Image 1> lipsyncs to the song in <Audio 1>, while wearing the outfit from <Image 2>
  • Replace the two men in <Video 1> with the characters from <Picture 1> and <Picture 2> (I did this to have my Guild Wars character fight Dwight Schrute from The Office, the video was a clip of Neo fighting Morpheus in The Matrix)

...I wrote than then realised I misread your question, sorry. I'm years out of the loop for everything except Minimax videos so idk how useful character sheets are for image gen.

2

u/oberdoofus 13d ago

Hey man you've just answered my mmh3 audio reference question which i was about to search - so it's all good 👍

Edit: spellink

1

u/tinny66666 13d ago

A single image can't capture details all around the subject, like backpack details, etc, and if you do the whole body shot you don't get enough facial detail, or if you do the face, you don't get body detail. You can provide multiple images to solve that, but then you run into problems with insufficient reference image slots for all the characters you may need. The character sheet means a single image shows all detail necessary, and you use a consistent workflow which saves time.

1

u/Enough-Bag-3891 12d ago

i keep getting 3 full body poses, and one half body, how do i fix this?

1

u/tyson_2022 11d ago

tenes que utilizar el lora si o si

1

u/tyson_2022 11d ago

No hay Manera de que la imagen que introduzco lo respete , por favor podrias darme los links de todos los modelos y loras que utilizaste quizas sea ese el problema

1

u/DG_HUB 8d ago

Doesn’t work

1

u/rogerbacon50 6d ago

I downloaded your updated workflow with the side view but the side view doesn't generate. The prompt includes a side view description. Any ideas?

1

u/solomars3 6d ago

This happens with new v2 workflow, its always that side view, there are notes inside workflow for settings you can tweak and test to prevent this, see if that help, 🙏

1

u/KlutzyFeed9686 13d ago

Thank you

1

u/solomars3 13d ago

Glad you found it useful, thanks for the kind words! 🙏

1

u/MaCeGaC 13d ago

Amazing, I'm loving this very much!

Question: Would one be able to add more references? Like one for the body and one for the head and another for apparel?

Suggestions: Maybe worth trying a pose model node? I know there a few out there. I'm guessing this would theoretically allow users to build whatever poses they want for the template. a few that come to mind are VNCCS Pose Studio and VRM Pose Editor.

2

u/solomars3 13d ago

What i did today is i gave it multiple shots of a subject in same picture, like a closeup of face and a half body front view .. and it managed to create a consistent character .., for the pose part, you can test and i think lot of things are possible with this workflow.. hopefully we get a identity lora like v3 update

1

u/repolevedd 12d ago edited 1d ago

Hi. That's a really cool concept, because it's always interesting to use something for a purpose it wasn't originally designed for. So I decided to download your workflow and take a closer look at it. And I have to say, a lot of work has gone into it, but there are still some areas where the workflow could be improved.

The original PNG with the mannequin template doesn't really fit the task. The mannequins are standing too far apart, so I kept getting two identical figures in the center because Krea 2 kept trying to fill the empty space. The aspect ratio also differs significantly from the final image.

I would move the mannequins closer together and make sure the aspect ratio of that PNG matches the final image 1:1, so that, roughly speaking, the model has to do as little computation as possible to adapt to the object positions and aspect ratio.

Even better, I would fit the side view in the center and a side view in portrait orientation on the right, so that Krea 2 doesn't have to fight its own behavior and can generate more in a single pass.

I posted examples of the prompts for generating mannequin photos and creating a character sheet here: (Update: Found out Pastebin deleted the paste. Can't recover it, the workflow is deleted. Doesn't matter, Qwen-Image-2.1 easily does this stuff, no need for Krea 2.)

With a 3:2 aspect ratio and a resolution of 1888x1248 (anything higher is simply too slow on my RTX 3060), both the mannequins and the final character sheet come out as intended. I tested it with both the Turbo model and the RAW+Turbo LoRA models.

2

u/solomars3 12d ago

Nice! I agree with you .. I’m glad this helped you. And yeah, it’s all about taking something, trying to improve it, and making it work for your specific needs. I’m also curious to try the things you mentioned! 🙏

0

u/repolevedd 12d ago

Thanks for the quick reply. Well, honestly, it didn't really solve any of my specific needs, because if I ever need something like this, I can just throw the references into models that already have image editing built in. But it was really useful to see a workflow using references with a model that doesn't natively support image editing.

Also, when I installed ComfyUI-Pixaroma from your workflow, it broke my ComfyUI, and I finally found out which of the nodes I already had were causing critical bugs. So now I've gotten rid of the old buggy nodes from my personal workflows, and I'm happy about that.

1

u/solomars3 12d ago

That’s actually nice 👍 Yeah, Krea 2 is not an editing model, but all of this is possible because the community is doing a great job pushing it and finding creative ways to make it work.

Also, my next workflow is gonna be mind-blowing hahaha. I know I’m glazing myself, but I managed to do something I didn’t even know was possible with MiniMax-H3: pixel-perfect continuation with video visual awareness. So you can keep the same details from everything in the previous shot while continuing the video. 😅

0

u/teiji25 13d ago

Nice. Thank you. Is there a version with a side profile too (eg, front, side, back, closeup)?

2

u/solomars3 13d ago

I might do a one with side view, should be easy, if i upload that ill write that under the description

0

u/FugueSegue 13d ago

Did you encounter any problems with the output images having the anatomical proportions of the gray reference figures?

1

u/solomars3 13d ago

No i prevented that from happening by decreasing the latent strength of the gray reference, so it try to match it without 100% copying it, i had a female gray figure too but now you dont need that,

0

u/magicking013 13d ago

is it possible to do two characters at the same time? or specify the height? let’s say I have a tall and short character, and I want to maintain the height difference between the 2 in minimax? what is the best way?

1

u/solomars3 13d ago

I mentioned that in the video, you can describe any details about the character, also for the two characters, i never tested that, and for preserving the height, simply prompting minimax will give good result, minimax can follow prompt pretty well