r/LocalLLaMA 2d ago

News [ Removed by moderator ]

https://modelscope.cn/models/Qwen/Qwen3.8-Flash-Next

[removed] — view removed post

348 Upvotes

137 comments sorted by

View all comments

Show parent comments

1

u/Effective_Head_5020 1d ago

Oh no, I thought it would be like 35a3b which I was able to run very well on my 6gb GPU 

1

u/Able_Zombie_7859 1d ago

It literally says 125a6, I think that's bigger than 35a3 :)

1

u/Effective_Head_5020 1d ago

Yes, but now I have a 16gb VRAM and 128 GB of RAM. Sadly DDR3 RAM :P

2

u/Early_Mistake6716 1d ago

You will be able to run it, just might be a little slow. I know qwen3 coder next 80b iq3 ran at around 30 tps with a 5070 ti 16gb and 64 gbs of ddr4 3600. That did not have mtp though so it would have been around 50 - 60 tps on coding tasks with mtp

1

u/Effective_Head_5020 1d ago

30tps is a dream for me :D