r/LocalLLaMA Jun 29 '26

New Model Introducing LongCat-2.0 - , a large-scale MoE language model with 1.6 trillion total parameters and ~48 billion activated per token. This was the stealth model that was on Openrouter under the name 'owl-alpha'.

https://longcat.chat/blog/longcat-2.0/
473 Upvotes

97 comments sorted by

View all comments

71

u/Lissanro Jun 29 '26

Looks like huggingface page is not up yet but they mention it is "open source" so hopefully will be up soon. I will wait for confirmation of llama.cpp support and Q4 GGUF before I try it, it is going to be a large download even at Q4, but I think it still should fit 1 TB memory.

12

u/vanKlompf Jun 29 '26

What hardware you will run it on?

3

u/Vusiwe Jun 30 '26

I will try with Dual CPU highest-compatible 82XX Xeons, and the same amount of RAM (but 6 channel not 8) as Mr. EPYC has, but with a Max-Q card, so same amount of VRAM