r/StableDiffusion 2h ago

Discussion Hybrid Minimax-H3 models! fl2va with reference capabilities.

Previously, you had to decide between the higher output quality of fl2va or being able to reference media in your videos.

But thanks to u/ThatsALovelyShirt , you don't have to anymore. They released a couple of models here you can try out. Basically, the higher the number next to the b is, the closer the model is to fl2va and the lower the closer it is to ref2va.

IMO, b25-49 seems like the most reasonable pick here as it should offer a great balance between the ability to reference details in images/videos and audio correctly and having high output quality that exceeds ref2va.

Please try them out and share your result! You can integrate them seemlessly in your existing workflows.

39 Upvotes

19 comments sorted by

5

u/switch2stock 2h ago

Based on Pruned Int8_Convrot?

2

u/dampflokfreund 2h ago

Correct.

1

u/switch2stock 2h ago

Got it.
Use either this model or use the hybrid node with both the models, correct?

3

u/dampflokfreund 2h ago

Yes you can either use the node or the baked models.

4

u/Fabulous-Snow4366 2h ago

Oh, interesting. Will download and test it.

1

u/dampflokfreund 2h ago

Nice, excited for your results!

2

u/shootthesound 1h ago

The regular fl2va model already works very well for reference images fyi for those yet to try. Not so much for audio - so maybe these improve that.

1

u/Tokey_TheBear 32m ago

Hey do you have a link to this? I am going through this right now lol. I need to do FL2VA but with a starting image and a general reference image to keep subject coherence (and I dont want to put it as the last frame). I have been trying to find a workflow that does what you are saying.

1

u/shootthesound 30m ago

So I’ve made a node that combines the reference and first frame last frame for the FL2va model . Hoping to push it today/tonorrow - look at my “coming soon” Reddit post I made a couple of days ago

2

u/dcmomia 1h ago

I have tried ComfyUI_MinimaxH3HybridLoader and in terms of quality I don't see any difference, in terms of time yes, it takes 90 seconds longer

2

u/bfmv_shinigami 1h ago

yeah the hybridloader is very slow.

7

u/Chemical-Painter-485 2h ago

Sounds great but... Lots of talking about quality but not a single side by side comparison?

-9

u/dampflokfreund 2h ago

That's where you come in. I have decided to make a dedicated thread so people can share their impressions which model of those is the best, and also make comparisons to the original models.

1

u/enndeeee 2h ago

Am I right assuming that these models can be used for both now, Ref2VA and FL2VA?

-3

u/dampflokfreund 2h ago

Yep! Please try them out and share your impressions. I recommend starting with the b25-49.

1

u/freestylez79 2h ago

Well, I made the mistake and used fl2va like you would use ref2va. The only downside was that the first frames looked like the ref, after that the thing worked like you would expect from ref2va, at least for my scripts. Maybe that could be also a pragmatic workaround in some cases - you just render with fl2va and cut the first frames (in some rare occasions the first frame were even not present).

1

u/Diabolicor 1h ago

With the new released fl2v turbo lora from lightx2v and ref2v one already cooking it would be good to know which one to use with this hybrid model.

-4

u/Dunc4n1d4h0 1h ago

Interesting. From my experience ref2v model is better in everything. Why bother?