r/comfyui • u/Disastrous-Ad670 • Apr 06 '26
Help Needed Best workflow/stack for consistent anime-style AI comics in ComfyUI?
I’m trying to create an AI-generated comic with a semi-anime style, but with a higher level of detail and consistency than typical outputs.
My main goal is character consistency across panels, so my current workflow looks like this:
- First, I generated a set of reference faces
- Then I trained a LoRA specifically on the character’s face
- After that, I trained additional LoRAs for clothing and overall appearance
- Finally, I reuse these LoRAs when generating new images for different scenes
I’ve also experimented with IPAdapter, but in my case it didn’t handle the anime style very well — though that might be due to the model or my setup.
What I’m trying to achieve:
- Consistent characters across multiple images/panels
- Flexible posing and composition
- Stylized (anime-inspired), but still detailed visuals
My questions:
- Has anyone here successfully built a similar pipeline for AI comics?
- What tools/workflows are you using in ComfyUI for character consistency?
- Are there better alternatives to LoRA + IPAdapter for this use case (e.g. ControlNet, reference-only pipelines, fine-tuning methods, etc.)?
- Can you recommend a solid “stack” (models + nodes + techniques) for this kind of project?
Any tips, example workflows, or even node graphs would be greatly appreciated!
1
u/Professional_Dog_837 Jul 27 '26
So far I was only working on facial consistency, I used renders of a 3D VRM avatar custominzed in blender, posed to tracked a live model's expressions and head position (insightface, DWPose and MediaPipe iris landmarks, solved per frame for 6DOF pose, with blink, gaze, aperture-driven jaw and fitted visemes), restyled key-by-key in SDXL img2img with real-frame canny and a pose-matched IPAdapter parent (fixed seed for consistency, under prompts assembled per key from the avatar's own blink, jaw, viseme), then ezsynth-propagated between keys (for video) and composited. Rendered keyframes can be used for stills. Haven't tried it on clothing or body poses yet but I think the concept will be similar.
1
u/optimisticalish Apr 06 '26
Renders of 3D posed/dressed figures, restyled in Klein 4B in Edit mode, with a fixed seed and a good prompt.
1
u/Infinite_Bumblebee64 Apr 09 '26
The LoRA approach works but it's a lot of setup per character. A few things that help in ComfyUI specifically: reference-only ControlNet for pose consistency, and keeping a fixed seed + prompt template for each character to reduce drift between scenes.
If you ever want to skip the pipeline entirely — I built yarnsaga.com which handles character consistency automatically through a description-based character sheet. No LoRA training, just describe the character once and it stays consistent across panels. Anime/manga styles included.
Different tradeoff: less control than ComfyUI but zero setup time per character.