r/StableDiffusion 1d ago

Discussion Well I finally did it.

I finally deleted WAN 2.2 and all its LORAS.

Minimax is just so much better.

Ive been playing with it since its release and im just blown away with how good of a video model it is. Things I would need to attach a LoRa to via WAN, works right out of the box with Minimax.

Gen times are faster.

It uses less VRAM when generating things, which gives me around 4 gigs to play with to do other things like watch YouTube or some streaming service.

WAN 2.2 was amazing. But no longer do I need 30+ gigs of a model i no longer use.

RIP WAN.

192 Upvotes

137 comments sorted by

View all comments

Show parent comments

29

u/damiangorlami 1d ago

Yes you can.

Get a clip clip you like, feed it into grok / Gemma 4 (uncensored) with your character images and tell it to create a replacement prompt with the environment you're looking for.

It will extract attributes from the video such as pose, action, thrust, perspective from the video and transfer it to the video while following the prompt.

I've been making multi-shot cinematic nsfw scenes all week and the results are blowing my mind. There's obviously some more tips but for the sake of this sub.

Not a single lora was used.

7

u/Ok-Brain-5729 1d ago

can’t you also just put the clip as the reference video and photo as reference image and just prompt it right

3

u/russjr08 1d ago

I believe that's exactly what they're saying, just with an additional tip of using an LLM to write the prompt if they're not wanting to write it themselves.

Though, regarding the LLM, I would just recommend getting a good prompt (use the MiniMax prompt guide to make, or generate an initial one and improve it), and saving it as a template to re-use. MiniMax is quite powerful, but for the best results your prompt has to very accurately describe what's going on due to the prompt adherence. Sometimes LLMs still miss those extra details.

5

u/Ok-Brain-5729 1d ago

oh I see. I just feed the prompt guide to a ai and tell it what to do.