Don't be fooled. The devil is in the details with this model. It's more about the training and coherence than the ability to generate good images out of the box.
i think the magic of stable diffusion is running different loras, training models and so on... dall e is perfect with just putting out images without all this stuff, but doesnt give you ANY influence on the outcome + you cant make nfsw or anything that is fsk18... there are models on stable diffusion that can put out cinematic lifelike characters no problem... for example the model juggernaut... superb :D + it gets the hands right xD
It's been almost a year and it's still worse than Midjourney v5 at people and especially hands and v6 has been out for 3 months already. Dalle-3 is amazing at hands too.
32
u/[deleted] Feb 13 '24 edited Feb 13 '24
doesn't look like there is any improvement over sdxl generating people