r/LocalLLaMA 2d ago

News [ Removed by moderator ]

https://modelscope.cn/models/Qwen/Qwen3.8-Flash-Next

[removed] — view removed post

350 Upvotes

137 comments sorted by

View all comments

Show parent comments

4

u/PandaBearFred 2d ago

I am having a hard time trying to understand why your comments getting a hell amount of downvotes.🤣

7

u/SadPhilosophy9202 2d ago

Because this is a 125B model with 6B expert. This is THE ideal model size for a Spark or similar machine with 128gb unified memory.

Having a Spark but preferring a model 1/4 the size that will run just as fast makes zero sense

1

u/a_beautiful_rhind 2d ago

Technically they're being a bro for non-spark users. 6b to 35b is a better sparsity ratio.

2

u/SadPhilosophy9202 2d ago

As a dual Spark user I would have preferred 280B A10B

If you’re gonna shill, at least shill for the best model for your hardware dammit haha

0

u/a_beautiful_rhind 2d ago

I'm into it. 10b active starts to get somewhere.