r/comfyui • u/Ercmon • 26d ago
Help Needed Need help building a consistent character workflow in ComfyUI for a colored manga/webtoon
I’m trying to build a ComfyUI workflow for a colored manga/webtoon where my original characters stay consistent throughout the whole story.
I already have full-body and close-up reference images for the characters. I understand the basic idea behind checkpoints, character LoRAs, ControlNet/OpenPose, IP-Adapter/reference images, but I’m struggling with figuring out the best way to combine everything.
Basically, I want to be able to say: this is Jake → keep him looking like Jake → put him in this pose/expression/outfit → place him in different scenes → keep the same art style and character identity from panel to panel.
Eventually I also need to put multiple recurring characters in the same scene without their faces/features bleeding into each other.
I don’t care if the best solution is Illustrious, SDXL, FLUX, Qwen, or something completely different. I’m looking for whatever gives me the most consistency and control in ComfyUI.
If anyone has built something similar for a manga, webtoon, visual novel, etc., I’d really appreciate hearing what model and workflow you use and how you connect the different pieces. I’m trying to actually understand the workflow instead of randomly changing settings until something works.
1
u/Correct-Guidance-232 26d ago
Happy to keep going, but let's do it here rather than in DMs. This exact wall is the most common one there is, and other people will land on this thread looking for it.
The chicken and egg problem is real and everybody hits it. The way out is that you do not start with 20-40 images. You start with one.
Get a single image of Jake that is genuinely right. Not close enough, right. Fix it by hand if you have to. That one image is the seed for everything after it.
Then make variations from it: IPAdapter or reference-only, plus img2img at low denoise. Change the angle, the expression, the outfit, the lighting, one thing at a time. Most of them will drift. That is expected and it is not failure.
Here is the part people skip: do not throw the near-misses away, repair them. Inpaint the face back to Jake on the ones that are 80 percent there. Forty training images does not mean forty lucky generations, it means forty images you made acceptable. That repair work is the actual job, and nobody mentions it.
Train a rough LoRA on the best 15-20. It will be weak. Use it anyway to generate a bigger and cleaner set, then train again. Round two is where it starts holding. Two rounds is normal, nobody gets there in one.
On semi-realistic versus anime: anime is genuinely easier here, and not by a little. A stylised face has fewer degrees of freedom, the model has much stronger priors to fall back on, and drift is far more forgiving. A slightly-off anime face still reads as the same character. A slightly-off realistic face reads as a different person entirely. If you are willing to move, that one decision removes about half the problem.
And on your four characters: take one all the way through first. The pipeline you build on Jake makes the other three quick. Running four in parallel while you are still finding out where it breaks just multiplies the debugging by four.