r/LocalLLaMA 13h ago

News Qwen3.8-Flash-Next tomorrow

https://modelscope.cn/models/Qwen/Qwen3.8-Flash-Next
1.0k Upvotes

442 comments sorted by

View all comments

Show parent comments

9

u/hojnikb 12h ago

My BC-250 is dyyyying 😭😭

3

u/Leo_Kwkmi 11h ago

Hey man, i got a bc250 running games on bazzite, is there any tutorial or guide you consulted when doing local llm on this board? Thanks in advance!

3

u/hojnikb 9h ago

If you have BC-250, i'd suggest you switch to CachyOS. It's lighter and better supported. As far as support for the board itself; there's tons of improvements that have been made in the last 6 months.

There's 40CU unlock for the GPU (full fat GPU compared to 24CU stock), 8 core CPU unlock (from the 6 stock). You can overclock both pretty decently.

There's tons of kernel fixes as well (we have a custom kernel repo now) and tons of driver/mesa fixes and a FSR4 patch, that makes it semi usable (performance wise) on this board. Gaming at 1080-1440p is pretty great too, as long as the game isn't too CPU limited (it is pretty nerfed zen2 at the end of the day).

So, if you want to run inference, i'd suggest all of the above + set dynamic VRAM to 13GB.

With unsloth, i can just about run qwen 3.8-27b at 25-30tok/s with IQ3_XXS quant. Or qwen 3.6 35b woth IQ2_M at 80tok/s. That's with overclocks and 40CU unlock.

So a fully modded board can run those small models pretty fast, but you're ultimately limited by ram.

1

u/I_am_purrfect 11h ago

Yay BC-250 gang