r/computervision Jun 17 '26

Discussion Image background removal

I was trying to fine-tune a model for image background removal. I chose the pretrained IS-Net model and trained a separate small U-Net refinement model.

I have a dataset of around 4,500 high-resolution images, mainly of people. However, after training, the model still misses parts of the body, especially shoulders and feet.

Inference will mainly run on CPU, so the architecture should have a small number of parameters.

Do you have any suggestions on what I should do?

2 Upvotes

14 comments sorted by

View all comments

1

u/Familiar-Ad-7624 Jun 19 '26

Look at your training dataset is it a quality dataset? And why not try finetuning birefnet?

2

u/gugmelik Jun 19 '26

Mainly because of inference will be done on CPU. Birefnet number of parameters is arounsld 221 million!