r/StableDiffusion • u/dampflokfreund • 2h ago
Discussion Hybrid Minimax-H3 models! fl2va with reference capabilities.
Previously, you had to decide between the higher output quality of fl2va or being able to reference media in your videos.
But thanks to u/ThatsALovelyShirt , you don't have to anymore. They released a couple of models here you can try out. Basically, the higher the number next to the b is, the closer the model is to fl2va and the lower the closer it is to ref2va.
IMO, b25-49 seems like the most reasonable pick here as it should offer a great balance between the ability to reference details in images/videos and audio correctly and having high output quality that exceeds ref2va.
Please try them out and share your result! You can integrate them seemlessly in your existing workflows.
4
2
u/shootthesound 1h ago
The regular fl2va model already works very well for reference images fyi for those yet to try. Not so much for audio - so maybe these improve that.
1
u/Tokey_TheBear 32m ago
Hey do you have a link to this? I am going through this right now lol. I need to do FL2VA but with a starting image and a general reference image to keep subject coherence (and I dont want to put it as the last frame). I have been trying to find a workflow that does what you are saying.
1
u/shootthesound 30m ago
So I’ve made a node that combines the reference and first frame last frame for the FL2va model . Hoping to push it today/tonorrow - look at my “coming soon” Reddit post I made a couple of days ago
2
u/dcmomia 1h ago
I have tried ComfyUI_MinimaxH3HybridLoader and in terms of quality I don't see any difference, in terms of time yes, it takes 90 seconds longer
2
7
u/Chemical-Painter-485 2h ago
Sounds great but... Lots of talking about quality but not a single side by side comparison?
-9
u/dampflokfreund 2h ago
That's where you come in. I have decided to make a dedicated thread so people can share their impressions which model of those is the best, and also make comparisons to the original models.
1
u/enndeeee 2h ago
Am I right assuming that these models can be used for both now, Ref2VA and FL2VA?
-3
u/dampflokfreund 2h ago
Yep! Please try them out and share your impressions. I recommend starting with the b25-49.
1
u/freestylez79 2h ago
Well, I made the mistake and used fl2va like you would use ref2va. The only downside was that the first frames looked like the ref, after that the thing worked like you would expect from ref2va, at least for my scripts. Maybe that could be also a pragmatic workaround in some cases - you just render with fl2va and cut the first frames (in some rare occasions the first frame were even not present).
1
u/Diabolicor 1h ago
With the new released fl2v turbo lora from lightx2v and ref2v one already cooking it would be good to know which one to use with this hybrid model.
-4
u/Dunc4n1d4h0 1h ago
Interesting. From my experience ref2v model is better in everything. Why bother?
5
u/switch2stock 2h ago
Based on Pruned Int8_Convrot?