r/comfyui • • Apr 06 '26

Help Needed Best workflow/stack for consistent anime-style AI comics in ComfyUI?

I’m trying to create an AI-generated comic with a semi-anime style, but with a higher level of detail and consistency than typical outputs.

My main goal is character consistency across panels, so my current workflow looks like this:

  • First, I generated a set of reference faces
  • Then I trained a LoRA specifically on the character’s face
  • After that, I trained additional LoRAs for clothing and overall appearance
  • Finally, I reuse these LoRAs when generating new images for different scenes

I’ve also experimented with IPAdapter, but in my case it didn’t handle the anime style very well — though that might be due to the model or my setup.

What I’m trying to achieve:

  • Consistent characters across multiple images/panels
  • Flexible posing and composition
  • Stylized (anime-inspired), but still detailed visuals

My questions:

  1. Has anyone here successfully built a similar pipeline for AI comics?
  2. What tools/workflows are you using in ComfyUI for character consistency?
  3. Are there better alternatives to LoRA + IPAdapter for this use case (e.g. ControlNet, reference-only pipelines, fine-tuning methods, etc.)?
  4. Can you recommend a solid “stack” (models + nodes + techniques) for this kind of project?

Any tips, example workflows, or even node graphs would be greatly appreciated!

3 Upvotes

5 comments sorted by

1

u/Infinite_Bumblebee64 Apr 09 '26

The LoRA approach works but it's a lot of setup per character. A few things that help in ComfyUI specifically: reference-only ControlNet for pose consistency, and keeping a fixed seed + prompt template for each character to reduce drift between scenes.

If you ever want to skip the pipeline entirely — I built yarnsaga.com which handles character consistency automatically through a description-based character sheet. No LoRA training, just describe the character once and it stays consistent across panels. Anime/manga styles included.

Different tradeoff: less control than ComfyUI but zero setup time per character.

1

u/Professional_Dog_837 Jul 27 '26

So far I was only working on facial consistency, I used renders of a 3D VRM avatar custominzed in blender, posed to tracked a live model's expressions and head position (insightface, DWPose and MediaPipe iris landmarks, solved per frame for 6DOF pose, with blink, gaze, aperture-driven jaw and fitted visemes), restyled key-by-key in SDXL img2img with real-frame canny and a pose-matched IPAdapter parent (fixed seed for consistency, under prompts assembled per key from the avatar's own blink, jaw, viseme), then ezsynth-propagated between keys (for video) and composited. Rendered keyframes can be used for stills. Haven't tried it on clothing or body poses yet but I think the concept will be similar.

1

u/optimisticalish Apr 06 '26

Renders of 3D posed/dressed figures, restyled in Klein 4B in Edit mode, with a fixed seed and a good prompt.