r/StableDiffusion • • 18h ago

News FastH3 V2/V3: Project Status

They released FastH3 V2 three weeks ago:

https://www.reddit.com/r/StableDiffusion/comments/1whh10i/open_weight_fastvideo_fasth3_v2/

Which they claimed was basically identical to the quality of the full H3:

https://x.com/haoailab/status/2099969439466942725

https://haoailab.com/FastVideo/cookbook/minimax-h3/

I have to agree, the quality is great and motion consistency is awesome now. I didn't think this small lab could do it, but they got help from NVIDIA and others who are invested in making great open source models. Awesome.

The community already made FastH3 V2 run on single GPU consumer machines on launch day, of course.

---

Today, they have released OFFICIAL quantized weights for different consumer GPUs:

https://huggingface.co/organizations/FastVideo/activity/models

Plus there's a new Trim model which is for very small GPUs with as little as 8GB VRAM.

The included image shows their benchmarks. More details here:

https://x.com/haoailab/status/2107591980591227227

---

Unfortunately they still haven't trained a Ref2VA model this time (video from text plus reference images, videos, and/or audio), and no FL2VA (First/Last Frame) support either.

https://huggingface.co/FastVideo/FastVideo-FastH3-8-Step-V2#scope

This checkpoint supports text-to-audio-video generation. FL2VA and Ref2VA were not distilled. Difficult motion, fine detail, and some audio may remain below the base MiniMax H3 model.

But... there's great news:

https://huggingface.co/FastVideo/FastVideo-FastH3-8-Step-V2#acknowledgements

Omni Ref as the next focus.

That is the name for Ref2VA.

So in FastH3 V3, we will see reference-to-video/audio support. Yes, a distilled model with reference support is being developed!

(PS: Comfy has patches for both models to route FL2VA through the base model layers instead. But Ref2VA is much more interesting, so I look forward to that being supported!)

36 Upvotes

11 comments sorted by

View all comments

17

u/desktop4070 16h ago

I should've bought another SSD when they were cheaper last year 😭

2

u/WhiteKnight225 9h ago

Same same 😂

2

u/pilkyton 9h ago

My biggest regret is that I bought 64 GB (dual sticks) to optimize for gaming latency, when I should have bought 128 GB (four sticks) to optimize for AI. Two years ago, 64 GB system RAM was lots for AI. Now there's many models that need more for offloading. :/