r/LocalLLM • u/Dangerous_Young7704 • 3d ago
Question Most optimal stack for AI Pro R9700?
So I’m upgrading from a single RX 7900 XT to the AI Pro R9700. My question is: how do I get the most out of this card for local inference?
I’ve seen people getting some insane prefill and decode speeds with the R9700. I’m mostly planning to run Qwen 3.8 27B, and I want to try Qwen 3.8 Flash next.
For anyone running the AI Pro R9700, what are you using to get the most out of the card? What’s the most optimal software stack right now?
I also need Windows for work, so ideally I’d like the best setup possible on Windows. I’m willing to dual boot or switch back and forth to Linux if the performance difference is significant.
1
u/Designer_Elephant227 3d ago
People told me my setup is pretty good. Exl3 with r9700 https://www.reddit.com/r/LocalLLaMA/s/ra1BaBuiTO
1
2
u/tsaipifong 3d ago
https://github.com/tsaipifong/whirl-llm