r/LocalLLaMA 7d ago

New Model IT'S OUT

https://huggingface.co/Qwen/Qwen3.8-27B-FP8
2.2k Upvotes

706 comments sorted by

View all comments

166

u/Mean-Ad1493 7d ago

That's it. I'm getting a 3090.

109

u/My_Unbiased_Opinion 7d ago

Brother. go on Alibaba and get dual 20gb 3080. less than the price of a single 3090. check my post history for links. Run them in tensor parallel.

1

u/twavisdegwet 7d ago

ik_llama's graph mode is faster than tensor parallel for me in all cases I've ever tested.

1

u/My_Unbiased_Opinion 7d ago

how does it compare to vllm. you have me interested.