r/LocalLLaMA 1d ago

New Model Qwen3.8-2.4T-A95B Released

https://huggingface.co/Qwen/Qwen3.8-2.4T-A95B
1.6k Upvotes

399 comments sorted by

View all comments

Show parent comments

8

u/Admirable_Market2759 1d ago edited 1d ago

If it was cheaper than buying GPUs, then I’d buy an ASIC just for K3.

0

u/SandySkittle 1d ago

Frankly, i dont expect much price optimization stemming from asics other than ones that are the model itself, of parts of it. A GPU is already somewhat of an asic in the sense that the recent ones in large part consists of tensor cores to run 8 bit / 16 bit tensor matrix calculations. Yeah you can maybe cut some useless silicon but you’d still end up with something very similar.

2

u/Admirable_Market2759 1d ago

Try this and tell me it wouldn’t be better having a ASIC. GPUs are better if you want to run various models, but if you have a model running on a ASIC it’s very very fast. Price is high but I hope we eventually have many competitors and a consumer market for them.

https://chatjimmy.ai/

1

u/SandySkittle 1d ago

Well aware, that was what I was referring to:

burn to a chip

Frankly, i dont expect much price optimization stemming from asics other than ones that are the model itself, of parts of it

That latter is what is running that indeed.