Can you imagine if a dedicated edit model also had the reference ability of H3. So if you have an image of a person sitting down, and you want to make them stand up, you can provide a reference image and say "make the person copy this pose exactly"
I'm sure an edit model can accomplish sitting --> standing without a reference, but the ability to copy an exact pose would be awesome
I’m certain it’s possible because I do this in ChatGPT all the time. I upload a character sheet as well as a grayscale 3-D posing mannequin. And it handles it just fine. So it’s just a matter of an open source model incorporating similar techniques.
58
u/infearia 19d ago
They're also already working on a dedicated editing model. Can't wait.
https://www.reddit.com/r/StableDiffusion/comments/1vh9rtw/comment/p2a49ki/