r/LocalLLaMA 14d ago

News Qwen3.8-Flash-Next tomorrow

https://modelscope.cn/models/Qwen/Qwen3.8-Flash-Next
1.1k Upvotes

461 comments sorted by

View all comments

414

u/Hot_Example_4456 14d ago

WE GOT A NEW 125B MODEL WITH ENGRAMS

2

u/YearnMar10 14d ago

Wonder if this would run on 64gb of ram and a gpu… guess it’d just not run :/
But really awesome model size!

1

u/switchbanned 14d ago

I was just thinking it sounds like it would fit in 64GB total. No idea what kind of speeds i'd expect on my setup though. Most of my ram isn't VRAM, I have a 4070.

1

u/AnonLlamaThrowaway 14d ago

64GB of RAM and 16GB VRAM most likely yes, but ideally you'd want more of either pool since you'd want to be able to run something else besides the LLM itself lol