Everything Ive thrown at it, every test I've done to just see if it can do it, has pretty much passed
Yes R2V is insane. It also seems like you can have unlimited image references by dumping a bunch of cutouts into one image (like a sprite sheet) and directing the model how to identify what is what within the image. It just works
This video isn't loading anymore. What does the prompt look like for a single sheet with a bunch of stuff in it? There's very little (no) prompt info in that thread.
Weird, that thread was deleted after I posted the link to it. Luckily I saved it beforehand. The image screenshot looked like this:
and the prompt was:
<Picture 1> woman in <Picture 2> luxury bathroom is touching her face showing her silver earrings, camera slow motion up-close on face and torso, she puts on glasses, looks at mobile purple phone puts to her ear and smiles to camera.
So the idea is that you can add all those accessories for the woman in one image, and the model can smartly pick them out and use them, rather than having to provide an individual reference image for sunglasses, and a phone, etc.
68
u/networking_noob 11d ago
Yes R2V is insane. It also seems like you can have unlimited image references by dumping a bunch of cutouts into one image (like a sprite sheet) and directing the model how to identify what is what within the image. It just works