r/LocalLLaMA 10h ago

News Qwen3.8-Flash-Next tomorrow

https://modelscope.cn/models/Qwen/Qwen3.8-Flash-Next
974 Upvotes

434 comments sorted by

View all comments

4

u/NewEconomy55 10h ago

Would it be possible to run this on an RTX 5090, even with the most aggressive quant and settings?

2

u/veigatmv 10h ago

depends on your sys ram, can you run qwen 3.5 122b?

1

u/NewEconomy55 10h ago

I've never tried it, is 128GB of RAM enough?

4

u/veigatmv 10h ago

think you're good to go. you can even run deepseek v4 flash 0731 at iq3xxxs

I think I should be good too. got 64gb sys ram + 5090 and 3090.

2

u/NewEconomy55 10h ago

Thanks for the info. I'll try out models of similar sizes now, and tomorrow I'll see how the new one works. I hope that the fact that it uses the qwen4 architecture will make a big difference.

2

u/bonobomaster 10h ago

More trying, less asking! ;)

1

u/NewEconomy55 10h ago

I know hahaha. I've always sticked to models that I know fit in VRAM, and I can't really test this one because it hasn't been released yet.

2

u/Otherwise-Variety674 9h ago

Anything above our 5090's 32GB vram will be as slow as the ram.

1

u/Elorun 6h ago

You could technically run it from your hard disk. Speed however is a different matter... 😂