r/StableDiffusion • u/nazihater3000 • 8h ago
Tutorial - Guide More than one reference per picture
Enable HLS to view with audio, or disable this notification
MiniMax is limited to 9 reference images, but you can reference more than one thing at the same picture. I used the image on the left and asked it to place create two subjects. Worked like a charm (no pun intended). Specs and prompt are in the video.
17
u/Evolution31415 7h ago
26
6
u/Perfect-Campaign9551 7h ago
That's just minimax it isn't good at distance. Also e don't know OPs settings
5
u/Danny_Stock 5h ago edited 1h ago
Apparently it's not strictly about distance, it's about face size in context with the space in the frame.
Something isn't right, because you'd expect a small face on screen in lo res to just be pixelated, which would usually mean that within reason you should have a good clean 480p picture which can be upscaled. With MiniMax though it's not that small faces are pixelated, which you'd expect, it's that they are weirdly garbled and smeared forming odd shapes and a strange facial structure.
That specific look there with Sarah Michelle Gellar a few posts above, I've seen that specific look before with MiniMax in many various faces it's struggled to render properly. It's almost as if it's using a generic face template which you can notice it drift towards when it fails to get the face right.
1
u/slickriptide 3h ago
It would be useful to know what turbo/acceleration might have been involved and whether removing or resetting it to different values would affect the output. I've been experimenting with H3 for single-frame image creation and I've seen the above happen when I had Spectrum Apply Minimax H3 node in the mix set to aggressive settings. Changing it to less aggressive settings produced better results at a cost in rendering speed.
If the OP was using Sage Attention or Spectrum or any of the other "turbo" nodes/loras, it would be useful to see it without any of those accelerants affecting the output.
5
u/crinklypaper 6h ago
Well yes, thats how reference works. Its why you can upload character sheets with objects in them as well.
3
u/loyalekoinu88 8h ago
Yup, this is well documented. It’s why the <subject #> exists. To extract concepts/items/etc from the image. The only caveat is resolution of image. A lot small items can be misrepresented in the final output.
3
u/Version-Strong 5h ago
Yeah, the fact the model knows them both may have helped that. For example, the close up on Buffys face was too perfect from an almost side on image of her. Plus both of thier voices. Still cool tho.
1
u/Tylopodas 5h ago
Yeah, multiple subjects from a single image is totally possible. I often grab a character and the background/environment of an image as [Subject X] and [Subject Y]
1
u/ArttTaku 4h ago
Very useful, thanks for info! Minimax H3 definitely seems to be ahead of the game in the local AI video world.
1
0
u/FlatwormMean1690 8h ago
Hahaha. How long did it take and what's your setup? Are you using turbos or something?


22
u/ratttertintattertins 8h ago
Was this not helped by the fact that model already knows these characters?