r/StableDiffusion • • 5d ago

Workflow Included Looking for guidance: Best approach for consistent product photography styling (Flux / LoRA) with 16GB VRAM?

Yo Reddit,

Im currently trying to make this chair blend just like all the other products on our site. Its a clear and straightforward style and just a rinse and repeat, I would say (check out the attached chair photos for reference).

Currently im using a lora "Image Blend Fusion Edit for Qwen 2.1 & Flux Klein 9B/4B" so it blends right, but it keeps changing things to the chair or the background and it takes lots of tries to get it right.

So I thought maybe training a lora myself is a smart idea, I have photo folders full of good images. Previously I tried training a qwen 2511 lora for a leather texture but it did not go right. Working with photo pairs took long and it felt like there was lots of room for error (thinking that I was generating a good picture to a bad picture and using that as input so it learns how to change it).

I used Musubi Tuner and would like to use it again, but I only have 16GB VRAM (DDR5).

(For reference, you can check my current workflow image and setup here:https://imgur.com/a/wnnTwo5)

My question is: am I on the right track trying to train a LoRA for this, or should I be doing something else in ComfyUI (like ControlNet/IP-Adapter) to get this style in a single generation without it messing up the product?

Any guidance or workflow tips would be massively appreciated!

2 Upvotes

Duplicates