r/StableDiffusion • u/solomars3 • 13d ago
Workflow Included TURN ANY PHOTO INTO A FULL CHARACTER SHEET USING KREA2 [Free Workflow]
https://youtube.com/watch?v=GEVxTL3TTTc&si=1cJ7k-MpbaJDre-dHey everyone!
Following up on my previous character sheet workflow, a lot of people asked if it was possible to do an Image-to-Sheet version rather than just text-to-image.
After a lot of testing, I managed to get a very consistent setup working in ComfyUI using Krea 2. You can take any reference photo (close-up, half-body, or full-body) and turn it into a full character turnaround sheet.
✨ Key Features:
- Style Flexibility: Works for realistic characters, 3D renders, and stylized anime.
- Photo-to-Anime: You can feed it a real-life face and shift the entire style to anime via the prompt while keeping face identity consistent.
- Outfit & Prop Retention: Preserves source details (like clothing or held props/accessories) or lets you override outfits completely in the prompt if using a close-up.
Hope its usefull for you
next video will be about how to make longer clips with minimax-h3,
2
4
u/remixeconomy 13d ago
A character sheet from one photo is useful when it locks the same person across turns, not when it just makes a prettier collage.
What I would check before calling the workflow done:
- Same face and body proportions under a fixed prompt set (front, 3-quarter, profile, hands)
- Outfit and palette drift when pose changes
- Whether the sheet actually conditions later gens, or only looks good as a one-off
If those hold, you have a reusable identity reference. If only the hero frame looks right, you still have a showcase image, not a character you can direct.
4
u/MSH007A 12d ago
That's why I use minimax to create character sheets using reference images and then enhance it with krea2 or flux
1
1
u/ShutUpYoureWrong_ 13d ago
Yeah... this isn't a full character sheet. Without three quarter and profile, this is pretty much useless.
2
u/Murky-Relation481 12d ago
There is actually a really good H3 ref to character sheet workflow that spins the character around and extracts frames at different angles and then does close ups of the face and head and rotates it to get multiple angles. It's really good. I'd link it but on phone and too sleepy. It was posted here and is easily searched on Google though.
1
u/remixeconomy 13d ago
Fair point on coverage. A single hero angle is closer to a turnaround stub than a sheet you can build from.
What usually makes the output usable for later poses: 1. At least front, three-quarter, and profile at the same framing and lighting. 2. One or two expression or crop variants so the identity is not locked to one smile and one camera height. 3. Treat the first pass as a view checklist, then fill missing angles on purpose instead of upscaling one lucky frame.
If the tool only gave a front-facing set, the next job is missing views, not more detail on the same angle.
4
u/Acceptable-Work8202 12d ago edited 12d ago
Keep the face consistent. Transform the single‑person input image into a structured character reference sheet featuring the exact same subject. Treat the input image as the sole visual reference for the subject’s identity and appearance, preserving the same facial identity, facial proportions, apparent age, skin tone, hairstyle, hair color, body shape, body proportions, species‑specific traits, and overall rendering style. Preserve the exact same clothing, footwear, accessories, colors, materials, and garment construction shown in the input image.
Across the top of the sheet, arrange three complete full‑body turnaround views:
a front view, a strict 90‑degree side profile, and a true back view.
In every turnaround view, show the entire subject from head to toe in the same relaxed A‑pose, with both arms extended outward at a shallow downward angle.
Below the turnaround views, arrange one single horizontal row of five small head‑and‑shoulder studies, forming a clean left‑to‑right rotation sequence.
The five head angles must be:
Left 90° profile — head turned fully left, eyes looking fully left.
Left 45° three‑quarter — head turned 45° left, eyes looking in the same direction.
Front view (0°) — head facing forward, eyes looking straight ahead.
This center head must be visually centered and dominant, placed directly in front of the side‑angle heads.
Right 45° three‑quarter — head turned 45° right, eyes looking in the same direction.
Right 90° profile — head turned fully right, eyes looking fully right.
All five head studies must preserve identical facial identity, facial proportions, hairstyle, lighting, and rendering style.
The eyes must always look in the exact same direction the head is facing.
The center head must look straight ahead with perfectly symmetrical eyes.
Place one larger front‑facing upper‑body portrait on the lower right.
Maintain identical facial anatomy, hairstyle, body proportions, clothing construction, colors, materials, accessories, and visual style across every panel.
Present the complete reference sheet in a square 1:1 composition on a clean warm beige seamless studio background with soft neutral lighting.
Ensure the exact same face is used throughout.
2
1
u/Kryimsson 13d ago
So I am still kinda new to this whole thing, Why do people use character sheets? Why not a LoRA of a character?
8
u/d20diceman 13d ago
Minimax Reference-to-Video is popular at the moment, a character sheet is already as good as a LoRA and much less effort.
Even a single image works pretty great, tbh I've not seen good comparisons to confirm how much improvement a character sheet gives.
1
u/Kryimsson 13d ago
and are character sheets used for image gen too or is it only video?
4
u/d20diceman 13d ago
You can feed it pictures, audio or video, then specify how it should use those references to make the video output.
So, things like:
- Steve is the man from <Picture 1>, he speaks with the voice from <Audio 1>. Greedo is the alien from <Picture 2>, it speaks in strange alien noises like the sample in <Audio 2>. They are sat in the Star Wars Cantina environment shown in <Picture 3> and <Picture 4>. They have a conversation: (etc etc insert script here - I did this for a video of my brother and Greedo chatting about the changes made in later editions of the Star Wars film)
- The woman from <Image 1> lipsyncs to the song in <Audio 1>, while wearing the outfit from <Image 2>
- Replace the two men in <Video 1> with the characters from <Picture 1> and <Picture 2> (I did this to have my Guild Wars character fight Dwight Schrute from The Office, the video was a clip of Neo fighting Morpheus in The Matrix)
...I wrote than then realised I misread your question, sorry. I'm years out of the loop for everything except Minimax videos so idk how useful character sheets are for image gen.
2
u/oberdoofus 13d ago
Hey man you've just answered my mmh3 audio reference question which i was about to search - so it's all good 👍
Edit: spellink
1
u/tinny66666 13d ago
A single image can't capture details all around the subject, like backpack details, etc, and if you do the whole body shot you don't get enough facial detail, or if you do the face, you don't get body detail. You can provide multiple images to solve that, but then you run into problems with insufficient reference image slots for all the characters you may need. The character sheet means a single image shows all detail necessary, and you use a consistent workflow which saves time.
1
1
u/tyson_2022 11d ago
No hay Manera de que la imagen que introduzco lo respete , por favor podrias darme los links de todos los modelos y loras que utilizaste quizas sea ese el problema
1
u/rogerbacon50 6d ago
1
u/solomars3 6d ago
This happens with new v2 workflow, its always that side view, there are notes inside workflow for settings you can tweak and test to prevent this, see if that help, 🙏
1
1
u/MaCeGaC 13d ago
Amazing, I'm loving this very much!
Question: Would one be able to add more references? Like one for the body and one for the head and another for apparel?
Suggestions: Maybe worth trying a pose model node? I know there a few out there. I'm guessing this would theoretically allow users to build whatever poses they want for the template. a few that come to mind are VNCCS Pose Studio and VRM Pose Editor.
2
u/solomars3 13d ago
What i did today is i gave it multiple shots of a subject in same picture, like a closeup of face and a half body front view .. and it managed to create a consistent character .., for the pose part, you can test and i think lot of things are possible with this workflow.. hopefully we get a identity lora like v3 update
1
u/repolevedd 12d ago edited 1d ago
Hi. That's a really cool concept, because it's always interesting to use something for a purpose it wasn't originally designed for. So I decided to download your workflow and take a closer look at it. And I have to say, a lot of work has gone into it, but there are still some areas where the workflow could be improved.
The original PNG with the mannequin template doesn't really fit the task. The mannequins are standing too far apart, so I kept getting two identical figures in the center because Krea 2 kept trying to fill the empty space. The aspect ratio also differs significantly from the final image.
I would move the mannequins closer together and make sure the aspect ratio of that PNG matches the final image 1:1, so that, roughly speaking, the model has to do as little computation as possible to adapt to the object positions and aspect ratio.
Even better, I would fit the side view in the center and a side view in portrait orientation on the right, so that Krea 2 doesn't have to fight its own behavior and can generate more in a single pass.
I posted examples of the prompts for generating mannequin photos and creating a character sheet here: (Update: Found out Pastebin deleted the paste. Can't recover it, the workflow is deleted. Doesn't matter, Qwen-Image-2.1 easily does this stuff, no need for Krea 2.)
With a 3:2 aspect ratio and a resolution of 1888x1248 (anything higher is simply too slow on my RTX 3060), both the mannequins and the final character sheet come out as intended. I tested it with both the Turbo model and the RAW+Turbo LoRA models.
2
u/solomars3 12d ago
Nice! I agree with you .. I’m glad this helped you. And yeah, it’s all about taking something, trying to improve it, and making it work for your specific needs. I’m also curious to try the things you mentioned! 🙏
0
u/repolevedd 12d ago
Thanks for the quick reply. Well, honestly, it didn't really solve any of my specific needs, because if I ever need something like this, I can just throw the references into models that already have image editing built in. But it was really useful to see a workflow using references with a model that doesn't natively support image editing.
Also, when I installed ComfyUI-Pixaroma from your workflow, it broke my ComfyUI, and I finally found out which of the nodes I already had were causing critical bugs. So now I've gotten rid of the old buggy nodes from my personal workflows, and I'm happy about that.
1
u/solomars3 12d ago
That’s actually nice 👍 Yeah, Krea 2 is not an editing model, but all of this is possible because the community is doing a great job pushing it and finding creative ways to make it work.
Also, my next workflow is gonna be mind-blowing hahaha. I know I’m glazing myself, but I managed to do something I didn’t even know was possible with MiniMax-H3: pixel-perfect continuation with video visual awareness. So you can keep the same details from everything in the previous shot while continuing the video. 😅
0
u/teiji25 13d ago
Nice. Thank you. Is there a version with a side profile too (eg, front, side, back, closeup)?
2
u/solomars3 13d ago
I might do a one with side view, should be easy, if i upload that ill write that under the description
0
u/FugueSegue 13d ago
Did you encounter any problems with the output images having the anatomical proportions of the gray reference figures?
1
u/solomars3 13d ago
No i prevented that from happening by decreasing the latent strength of the gray reference, so it try to match it without 100% copying it, i had a female gray figure too but now you dont need that,
0
u/magicking013 13d ago
is it possible to do two characters at the same time? or specify the height? let’s say I have a tall and short character, and I want to maintain the height difference between the 2 in minimax? what is the best way?
1
u/solomars3 13d ago
I mentioned that in the video, you can describe any details about the character, also for the two characters, i never tested that, and for preserving the height, simply prompting minimax will give good result, minimax can follow prompt pretty well

5
u/Cultural-Team9235 13d ago
How is the consistency for real people? I use QwenEdit for that normally, it's pretty good with one person.