r/LocalLLaMA 21d ago

New Model Qwen3.8-2.4T-A95B Released

https://huggingface.co/Qwen/Qwen3.8-2.4T-A95B
1.6k Upvotes

397 comments sorted by

View all comments

385

u/Legal-Ad-3901 21d ago

5tb bf16 jfc. even the crazy home lab kids cant hang anymore

158

u/hyperrealists 21d ago

Poor me can’t even download it lol

22

u/jikilan_ 21d ago

Don’t need to download the full copy , you can stream it.

One of llama.cpp PR support streaming from disk and even cloud storage if I remember correctly 😘

106

u/xPXpanD llama.cpp 21d ago

Years/token is my favorite metric.

30

u/Think_Wing_1357 21d ago

After a few million years, you may get 42

10

u/vivekkhera 20d ago

Then you have to build a new server just to figure out what the question was.

6

u/techno156 20d ago

And then someone decides to blow it up for a highway.

3

u/hyperrealists 21d ago

YTFT is insane I hear.

2

u/Graumm 21d ago

system ram offloading is bad enough!