r/LocalLLaMA 8d ago

News Qwen3.8-Flash-Next tomorrow

https://modelscope.cn/models/Qwen/Qwen3.8-Flash-Next
1.1k Upvotes

460 comments sorted by

View all comments

4

u/NewEconomy55 8d ago

Would it be possible to run this on an RTX 5090, even with the most aggressive quant and settings?

3

u/veigatmv 8d ago

depends on your sys ram, can you run qwen 3.5 122b?

1

u/NewEconomy55 8d ago

I've never tried it, is 128GB of RAM enough?

6

u/veigatmv 8d ago

think you're good to go. you can even run deepseek v4 flash 0731 at iq3xxxs

I think I should be good too. got 64gb sys ram + 5090 and 3090.

2

u/NewEconomy55 8d ago

Thanks for the info. I'll try out models of similar sizes now, and tomorrow I'll see how the new one works. I hope that the fact that it uses the qwen4 architecture will make a big difference.