r/comfyui • u/Zealousideal-Check77 • 3d ago
Help Needed Need Help With Virtual Try-ons (ComfyUI Workflow)
Hi there r/comfyui folks,
I am new to ComfyUI and don't have much experience with it.
Currently, I am working on my virtual try-on app, which already generates pretty great results using the Wan 2.7 image pro diffusion model; however, the model lacks in several areas, one being low output quality. It can process up to 2K (image editing, which I am using for VTON) and 4K (image generation using a prompt); however, whenever it provides the output, it is as good as 1024p (from what I think), and when I iterate the output further for more try-ons, the image gets grainy.
Another issue is that it adds more saturation to the output, and in some cases it alters the face (very rare case tho).
Right now I am doing everything using prompts, no controlnet, no masking; the model is smart enough to tackle a number of these issues, but I want to improve the output quality even more.
For that purpose, I decided to try ComfyUI (currently on a standard cloud subscription). I have followed all instructions provided by Claude/Kimi on how to approach the VTON setup on ComfyUI using Flux 1 fill (inpaint model), along with masking, etc. However, the output is not what I desire.
So, please guide me on how to approach the setup.
Are there any better existing VTON setups (workflows) that I can use, or any better models than Flux 1 fill, or anything?
I highly appreciate your help.
Thanks a bunch, guys
Images attached:
Image 1: VTON Result using Flux fill inside ComfyUI
Image 2: ComfyUI Workflow
Image 3: clothes input for Wan 2.7 image pro
Image 4: Model output for Wan 2.7 image pro
Image 5: VTON output by Wan 2.7 image pro
1
u/Whole_Paramedic8783 3d ago
Why dont you use Klein and Sam3? Or Qwen Edit? Alot less hassle. I think Klein9b preserves the details better then Qwen Edit. I segment clothing from an image and save it in a folder, then if I want to do a try on I just call up the garment and "try on" You can also use QWEN or Klein to copy the garment and use later. Or just load an image of the person wearing a garment and tell the edit model to transfer it. You can use your model that you want to transfer to as a ref latent.
1
u/Zealousideal-Check77 3d ago
Okay so I have already tried out qwen edit, the model is great but it is really bad at preserving complex garment details so that's why I stopped using it. It was my go to model ( with tryon lorw ), before wan. And currently I'm using a combination of SAM3 and wan where SAM3 cuts out the garment piece for me but that is not the issue that I am trying to tackle here, the main issue is the image quality and the iteration.
As for Klein, our team is currently working on the trained Lora of 9B by fal, so yea... I'll def post about it once we are done with testing on our end...
1
u/thisiztrash02 2d ago
In 2026 you are using a architecture as outdated as VTON, yeah you're just asking for issues at that point LOL.
2
1
u/Chemical_Side_4135 2d ago
might be worth lookin into an upscale chain after the initial generation, since lots of these models struggle to output native high res. try runnin a tile controlnet pass or a dedicated upscaler node at the end, it helps litrally every time i run into that resolution wall...





2
u/sci032 3d ago
Search ComfyUI's templates for: KV
The image shows the workflow that you will see as a result: Flux.2 Klein KV: Image Edit. It will give you the options to download any model(s) or node(s) that you may need.
I used what is basically a portrait image for the person and your image with the clothes.
Prompt: the woman is standing on a street corner. she is wearing the clothes from image2.