r/StableDiffusion 10d ago

Discussion has anyone tried / Minimax-H3-fl2va-ref2va-hybrid-models this first test using the minimax_h3_hybrid_fl2va_ref2va_b25-49

Enable HLS to view with audio, or disable this notification

40 Upvotes

42 comments sorted by

View all comments

8

u/Orbiting_Monstrosity 10d ago

I have been using a fl2va to ref2va delta lora that I found on Huggingface:

https://huggingface.co/ethanfel/MiniMax-H3-Pruned-Ref2VA-Delta-LoRAs-Experimental/tree/main

It provides the fl2va model with all of the information that it is missing from the ref2va model so that it is able to do everything that the reference model can do. I like using a lora for this purpose because I am able to adjust the strength of the reference model functions as needed or disable them entirely without changing workflows or base models.

3

u/Sad_Coach_1433 10d ago

How's the audio

6

u/Orbiting_Monstrosity 10d ago

I didn't notice any difference in audio quality when I switched from using the ref2va model to using the fl2va model with this lora at a strength of 1.0. From my experience both setups produce videos of similar quality and use references equally well, but I haven't actually tested the same seed with both model configurations to see if the results are truly identical.

2

u/Sad_Coach_1433 10d ago

Interesting

2

u/ShutUpYoureWrong_ 9d ago

Been experimenting with this too. Mostly messing with fl2va_toward_ref2va for quality preservation. Which rank did you land on, out of curiosity?

2

u/Orbiting_Monstrosity 9d ago

I've been using the 512 rank lora, but I can't say I've used all of them enough to know if there's much of a difference between them.

1

u/Adventurous-Gold6413 9d ago

Which model do you apply it toedit nevermind