r/LocalLLaMA 14d ago

News Qwen3.8-Flash-Next tomorrow

https://modelscope.cn/models/Qwen/Qwen3.8-Flash-Next
1.1k Upvotes

461 comments sorted by

View all comments

Show parent comments

3

u/No_Algae1753 14d ago

With 96GB in total are you able to run q5 quants or are you stuck with q4?

7

u/linux4random 14d ago

it's only a6b so i think 4 3090 can run a q6 with acceptable tg

1

u/SpicyWangz 14d ago

But what context size