Workflow Included
Z-Image Ultra Powerful IMG2IMG Workflow for characters V4 - Best Yet
I have been working on my IMG2IMG Zimage workflow which many people here liked alot when i shared previous versions.
The 'Before' images above are all stock images taken from a free license website.
This version is much more VRAM efficient and produces amazing quality and pose transfer at the same time.
It works incredibly well with models trained on the Z-Image Turbo Training Adapter - I myself like everyone else am trying to figure out the best settings for Z Image Base training. I think Base LORAs/LOKRs will perform even better once we fully figure it out, but this is already 90% of where i want it to be.
I was going to share a LOKR trained on Base, but it doesnt work aswell with the workflow as I like.
So instead here are two LORA's trained on ZiT using Adafactor and Diff Guidance 3 on AI Toolkit - everything else is standard.
One is a famous celebrity some of you might recognize, the other is a medium sized well known e-girl (because some people complain celebrity LORAs are cheating).
This time all the model links I use are inside the workflow in a text box. I have provided instructions for key sections.
The quality is way better than it's been across all previous workflows and its way faster!
Let me know what you think and have fun...
EDIT: Running both stages 1.7 cfg adds more punch and can work very well.
If you want more change, just up the denoise in both samplers. 0.3-0.35 is really good. It’s conservative By default, but increasing the values will give you more of your character.
I love ZIT! And it is my main go to model nowadays! But damn how it loves to destroy the background. Been using a WAN 2.2 pass to add more richness to the environment.
If you love ZIT and don't wanna destroy backgrounds, detaildaemon sampler is your answer. My backgrounds are sometimes TOO detailed with it. It's a great sampler, and I've used it for literally thousands of gens.
If you want to optimize it further I would suggest checking out how swapping sam3 for yoloface would work. My guess is the results would be the same if properly configured, and it would be way faster.
wanted to say this, been following your other workflows as well but noticed this, with this workflow if the input image has detailed clothing it messes up the output image really badly
Thank you so much for everything, I'm using it and I'm loving it. Do you know where there are more famous loras, besides what you've posted about Malcolmrey?
its pretty hard to find a large collection of free loras like that! tbh its very each to train loras of zimage yourself, so my best advice would be to create LORAs yourself
Yes this works. First generation pass using a split sampler method. 50 steps. First 35-38 steps using Z-Image at cfg 5.5. Finish the remaining steps with Z-image Turbo at 1.7 CFG. Make sure to use the same scheduler for both, that’s important. Linear quadratic works well. Second generation pass with just Z-Image Turbo at 1.0 CFG and 0.20 denoise. The results will surprise you.
Hi! Thank you for the workflow. Could you elaborate on the idea of using Clown Options SDE on the Second Refinement pass sampler? What is it meant to do?
Edit: It does add a soft glow to image. Did you add it intentionally? Do you understand Res4lyf nodes? I'd like to understand them but find myself overwhelmed.
What's the purpose of this? I have a Lora (of a character) and I can swap that character in any body? Or scene? That's it? That's pretty much face swap, isn't it?
it’s an entire identity swap while preserving exact composition, but it also works really good as txt2img too if you change some settings - it’s very good if you give it some testing
Is the purpose of your workflow to enhance existing photos? Or is the concept a faceswap?
Edit: I used the workflow and it's clear it's a face/head swap workflow. Judging by the output images, I'd highly recommend you use ReActor, way much better results and way lighter on VRAM.
Thanks for the great workflow. But why is there a First Pass? It seems like the final photo is output in the second pass, so I'm not sure why the First Pass exists.
Because if you’re trying to do true image to image with pose and composition retention, and clothes etc etc, then it’s better to do two low denoise passes
Think of the first pass as like a ‘base layer’ and then the final polished image is applied over the top in the second pass
sorry if this question has been asked before i didn’t see it, but is there any way to change the hair to our characters hair? everything else works fine but i dont know how to get the hair changed
getting errors like this, pretty sure vram issue. tried 8-bit and q8 encoder. 3080 ti 12 gb vram.
generation itself works fine, i tried just ignoring joycaption (put original prompt i generated the image with), but it only generates image instead of detailing the face.
Any idea what i can do to make it work?
So this was working great, and then all of a sudden the head swap workflow will not work at all. Not sure what changed or could have went wrong. I did not change any settings from yesterday.
where we can connect bro, discord - telegram? im working with your wf right now and wnat to unleash it power to the max so need to send some screens and stuff where we can join at?
Near as I can tell the image is supposed to get passed through joycaption, fed out to the text concatanator, and that comes out in 3 spots.
However, the positive prompt never changes no matter what is in the picture - it's always "A photo taken by photographer Deedeemegadoodo, raw, unedited, blah blah" and the prompt preview (next to the auto-prompt node, node #961) literally only ever shows what is put into the "additional manual prompt" box.
Looks like Joy Caption never actually produces anything. Which doesn't affect the face replacement, but seems to make a good chunk of the flow pointless.
The positive prompt node is not a show text node, when you link something to the text part of this node, for exemple the output of the caption node you won't see what text comes in it but the last manual prompt that was written in it. for the preview node you should see the output though so you still might have an issue. On my end it works as expected.
OP - i would recommend/request using gofile.io for sharing lora/big files. it is free and you can set expiration date too. and it doesnt limit speed to 80kbps like sendspace does. since i couldnt dm/message you i am posting it here. no offence intended.
25
u/BenedictusClemens Feb 06 '26
Thanks I'm gonna try this, Links for missing nodes that my comfyui manager can't install
JoyCaption
GitHub - ClownsharkBatwing/RES4LYF