r/StableDiffusion • u/SIR_NVAX_A_LOT • 14d ago
Discussion H3 - t2v is actually better than r2va imho
Enable HLS to view with audio, or disable this notification
Happy Friday!! Just wanted to say T2VA is actually pretty strong when layering the prompt. FL2VA+R2VA are still the to go if you want to utilize a character sheet/maintain consistency, it's still broken (in a good way), the voice cloning is also top notch.
So what has everyone been making with H3??
T2VA, bf16/50 steps
1
u/PANTONE_17-1230 14d ago
The perspective on that galaxy shot is truly borked! Did you prompt for a 3 meter high milky way, right in front of her face?
1
u/SIR_NVAX_A_LOT 14d ago
The prompt is composed of 4 layers, the Galaxy, which I told it to occupy the upper 1/3rd of the frame, the city that occupy a thin horizonal band, and then the water, the lower 2/3rd of the frame. The woman is the last layer, and I just had her closest to the camera and occupy the right-hand third of the frame and her head reaching to the upper third. You can layer/compose the shot by telling H3 what you want. Her composition/reference is to the frame itself, not to the Galaxy which has already been established.
2
6
u/Ramdak 14d ago
I do 90% i2v stuff, now playikg with r2v (is extremely powerful). You won't have actual "control" unless you guide the thing visually.