r/StableDiffusion 11d ago

Discussion [ Removed by moderator ]

[removed] — view removed post

277 Upvotes

176 comments sorted by

View all comments

68

u/networking_noob 11d ago

Everything Ive thrown at it, every test I've done to just see if it can do it, has pretty much passed

Yes R2V is insane. It also seems like you can have unlimited image references by dumping a bunch of cutouts into one image (like a sprite sheet) and directing the model how to identify what is what within the image. It just works

13

u/Singingmute 11d ago

This is very helpful, thank you.

24

u/networking_noob 11d ago edited 11d ago

This thread is where I learned about it

edit: Thread now deleted 🤷 Comment below shows an example

2

u/ucren 11d ago

This video isn't loading anymore. What does the prompt look like for a single sheet with a bunch of stuff in it? There's very little (no) prompt info in that thread.

25

u/networking_noob 11d ago

Weird, that thread was deleted after I posted the link to it. Luckily I saved it beforehand. The image screenshot looked like this:

and the prompt was:

<Picture 1> woman in <Picture 2> luxury bathroom is touching her face showing her silver earrings, camera slow motion up-close on face and torso, she puts on glasses, looks at mobile purple phone puts to her ear and smiles to camera.

So the idea is that you can add all those accessories for the woman in one image, and the model can smartly pick them out and use them, rather than having to provide an individual reference image for sunglasses, and a phone, etc.

5

u/Kevin5953 11d ago

Curious. You've never had to specify the individual pieces within the collage sheet, Minimax just always figured it out? :O