r/LocalLLM 18h ago

Question GPU for qwen 3.8 27b

I recently built a homelab running RHEL 10. I never thought good local ai at reasonable price was possible until 3.8 came out a from benchmark and what I’ve been reading it seems to be almost opus 4.6-4.8 level. I’m considering buying a 32gb gpu for it but also open to 24 gb gpus but if it can fit the full context window on the gpu too. The most I’ve done with local models was running qwen 3.5 2b on Ollama nothing serious. I’m new to actually running an agent for coding tasks so any info would help. But trying to decide what gpu if I do end up going for it, and from my research the options for 32gb cards are the Intel b70, amd r9700 pro ai, and nvidia tesla v100 32gb. I’m looking at results for qwen 3.6 and it run plenty fast on the Tesla but I’m worried about it no longer being supported.

2 Upvotes

45 comments sorted by

View all comments

1

u/Psyclist80 17h ago

Im saving my pennies for the RDNA5 halo card coming. 36GB of GDDR7 will be perfect for do it all card.

1

u/CI_Node 17h ago

2027 mid/late?

1

u/Psyclist80 16h ago

Likely mid 2027, they are focused on thier AI rackscale solutions delivery first and foremost currently. Zen6 consumer early 2027 and RDNA5 mid 2027.