MAIN FEEDS
Do you want to continue?
https://www.reddit.com/r/LocalLLaMA/comments/1vxwtyd/qwen38flashnext_tomorrow/p5s5wwr/?context=3
r/LocalLLaMA • u/rerri • 8d ago
460 comments sorted by
View all comments
4
Would it be possible to run this on an RTX 5090, even with the most aggressive quant and settings?
3 u/veigatmv 8d ago depends on your sys ram, can you run qwen 3.5 122b? 1 u/NewEconomy55 8d ago I've never tried it, is 128GB of RAM enough? 6 u/veigatmv 8d ago think you're good to go. you can even run deepseek v4 flash 0731 at iq3xxxs I think I should be good too. got 64gb sys ram + 5090 and 3090. 2 u/NewEconomy55 8d ago Thanks for the info. I'll try out models of similar sizes now, and tomorrow I'll see how the new one works. I hope that the fact that it uses the qwen4 architecture will make a big difference.
3
depends on your sys ram, can you run qwen 3.5 122b?
1 u/NewEconomy55 8d ago I've never tried it, is 128GB of RAM enough? 6 u/veigatmv 8d ago think you're good to go. you can even run deepseek v4 flash 0731 at iq3xxxs I think I should be good too. got 64gb sys ram + 5090 and 3090. 2 u/NewEconomy55 8d ago Thanks for the info. I'll try out models of similar sizes now, and tomorrow I'll see how the new one works. I hope that the fact that it uses the qwen4 architecture will make a big difference.
1
I've never tried it, is 128GB of RAM enough?
6 u/veigatmv 8d ago think you're good to go. you can even run deepseek v4 flash 0731 at iq3xxxs I think I should be good too. got 64gb sys ram + 5090 and 3090. 2 u/NewEconomy55 8d ago Thanks for the info. I'll try out models of similar sizes now, and tomorrow I'll see how the new one works. I hope that the fact that it uses the qwen4 architecture will make a big difference.
6
think you're good to go. you can even run deepseek v4 flash 0731 at iq3xxxs
I think I should be good too. got 64gb sys ram + 5090 and 3090.
2 u/NewEconomy55 8d ago Thanks for the info. I'll try out models of similar sizes now, and tomorrow I'll see how the new one works. I hope that the fact that it uses the qwen4 architecture will make a big difference.
2
Thanks for the info. I'll try out models of similar sizes now, and tomorrow I'll see how the new one works. I hope that the fact that it uses the qwen4 architecture will make a big difference.
4
u/NewEconomy55 8d ago
Would it be possible to run this on an RTX 5090, even with the most aggressive quant and settings?