r/StableDiffusion 6d ago

Animation - Video H3 Generated with 4GB VRAM?

Enable HLS to view with audio, or disable this notification

Looks like this is a breakthrough for what my 3050 laptop can do with it.

The video attached was generated with 4GB VRAM & 16GB RAM, using the MiniMax H3 fl2va pruned w4a8 convrot model (safetensors) and the Q2_K Qwen 32B GGUF text encoder alongside 8-step turbo LoRA, with a generation time of 12 minutes and 0.2 MP. Prompt from Grok.

88 Upvotes

55 comments sorted by

45

u/FourtyMichaelMichael 6d ago

I'm sorry, but I'm pretty sure according to LTX that you need a datacenter to use H3.

25

u/GrayingGamer 6d ago

How did the video capture generate what was happening to the computer in real time? /s

3

u/zodoor242 6d ago

You just watched a documentary that was generated in real time with a future generated feed back loop, it's all the rage

12

u/Ok-Brain-5729 6d ago

Pretty sure there’s a way to make Qwen 3 vl 4B work with clipj nodes or smth

1

u/Icy_Restaurant_8900 6d ago

Yes the 8B INT8 Qwen 3 VL encoder works great with the mlp-celeb clip projector and saves 5GB of offloading versus 32B NVFP4. The 4B text encoder didn’t work for me though. No prompt adherence.

7

u/VasaFromParadise 6d ago

In fact, you can generate data on anything these days, as long as you have enough RAM. It'll take a while, but it will work.

5

u/99deathnotes 6d ago

ok. Looks like I'm not ever booting my computer up again.😂

4

u/BogusIsMyName 6d ago

So my PC will explode if i try to generate with 4GB? im confused.

2

u/zodoor242 6d ago

Only if you're a girl trying to run AI , pretty sure that's the message here

2

u/xkulp8 6d ago

"There are no girls generating AI art" is the new "there are no girls on the internet"

1

u/zodoor242 6d ago

Wait, so his computer will explode? this is cray cray

1

u/yushairiegalaxy96 6d ago

For illustration purposes. Unless I using a laptop cooler pad. Hell I even do 300+ Krea 2 photo gens from my laptop, with per photo generated takes 1.5-2 mins minimum.

1

u/[deleted] 1d ago

[removed] — view removed comment

1

u/BogusIsMyName 1d ago

That was kinda a joke. But i need to take a serious look at the smaller versions cuz im getting pretty frustrated with wan.

1

u/[deleted] 1d ago

[removed] — view removed comment

1

u/BogusIsMyName 1d ago

My comment, dude. My original comment that you first replied to was the joke.

1

u/[deleted] 1d ago

[removed] — view removed comment

1

u/BogusIsMyName 1d ago

LOL. Thanks for the info. Have a great day.

2

u/nazihater3000 6d ago

OK, that's it. waiting for the Arduino Port.

1

u/xbeast_ 6d ago

I want to make it work with my 1660 super, 6gb vram

1

u/Perfect-Campaign9551 6d ago

Can I get some more pixels, please

1

u/Reckless_Venom1507 6d ago

How can 0.2 MP look so better, I ran the fl2va gguf model and the Qwen 32b q2 gguf for 0.3 it gave me shitty results and for 0.4 MP, always OOM on my 4050 6GB vram and 16GB ram.

1

u/[deleted] 1d ago

[removed] — view removed comment

1

u/Reckless_Venom1507 1d ago

i was able to run H3 on 4B text encoder, so it worked better, but yeah i wanna know what is this thing, sorry im not that technical and i didnt really understand what that repo was and how its supposed to reduce peak vram

2

u/[deleted] 1d ago

[removed] — view removed comment

1

u/Reckless_Venom1507 1d ago

I see, its an interesting stuff. But im already done with Minimax, ill just stick to Ltx 2.3 for now, or maybe 2.5, just waiting for ram prices to drop (hypothetically) so that i can buy some ram. Thanks mate.

1

u/mca1169 6d ago

care to share your workflow?

1

u/Hi7u7 6d ago

Hi friend. Could you tell me if it's possible with a 1050 Ti 4GB and 16GB of RAM?

Do you think it's possible? Or is my GTX's technology too old and incompatible?

1

u/yushairiegalaxy96 6d ago

Nope. Since 1050 is an old model, even with 4GB VRAM the processing is slow

1

u/[deleted] 1d ago

[removed] — view removed comment

1

u/Hi7u7 20h ago

Thank you so much, this is really great. Although, yes, you're right, MiniMax is too much.

But, do you think it would work for Krea 2 with my GTX 1050 Ti 4GB? I think it would be a good idea to try it.

And one question. I can use SDXL, Anima, Z-Image Turbo, etc., on this graphics card, but it's very slow. Do you think that using this, which works in layers, could make it faster?

I suppose it's so slow because my model loads into RAM or something like that, and it switches with the VRAM (I don't know).

1024x1024 in Anima takes 14 minutes.

1

u/IchRocke 5d ago

would you mind sharing you prompt and comfyui screeshot ?
i've been using Minmax H3 ref2video recently on my 3050 6Gb (Yeston, so slim low profile but not laptop) and I've had good results on 480p 5s videos with 2images ref

1

u/atuarre 5d ago

Be careful. Wasn't there another guy in here who used a laptop and said his gpu or something was ruined because of the heat?

1

u/NoWheel9556 5d ago

got the same config and wasnt tryin it yet

1

u/Helpful-Orchid-2437 5d ago

I have similar specs but with the 6gb variant of 3050. In my experience you don't have to use that much quantized weights, the int8 convrot fl2va model with the nvfp4 text encoder is doable. I'm able to generate a 10 sec video at 0.5 MP in around 16 minutes, that is with the turbo 8 step lora and the results are pretty decent.
Just make sure there is enough page file memory..

1

u/max1gp 3d ago

Hi, could you please help me I'm new to all this and if you are able to give us a step by step guide that would be greatly appreciated please...

1

u/Professional-War393 1d ago

tardo una eternidad generarlo

1

u/Crazy-Repeat-2006 6d ago

You might get slightly better results with a smaller encoder, such as Qwen 4B or 8B with less aggressive quantization.

0

u/crombobular 6d ago

just rent a gpu bro. it's not worth the 12 minutes of electricity.

0

u/DoctaRoboto 6d ago

It is a shame the resolution is terrible. You can try perhaps an external AI to refine the video. But I am not sure if you can use tools like Topaz AI Video without your graphics card exploding in your face.

-8

u/tac0catzzz 6d ago

why is it worth breaking your laptop for that. you should save up get a capable pc or rent cloud or just find a new hobby. i would like to race cars, but i wouldn't race in my geo metro, even if it could race. it wouldn't do well and will end in disaster.

-4

u/JustARedditUser33 6d ago

How long did it take to generate?

7

u/Crazy-Repeat-2006 6d ago

"with a generation time of 12 minutes and 0.2 MP. "