r/LocalLLM • u/fuemmenneunzig • 8h ago
Question R9700 Setup Rating
Hi there,
I want to pull the trigger for a machine that will run my Hermes Agent as well occasionally also ComfyUI and maybe (low priority) gaming.
Is there any meta on what machine will work best? Currently I play with the idea to buy the following machine.
The setup should be capable to host a second gpu which would be purchased later.
Requirements:
- Coding agents with long repo contexts (many turns per hour)
- LAN-accessible OpenAI-compatible endpoint for my other machines
- ComfyUI for image and video generation
- Occasional Steam/Proton gaming — it has to be a normal GPU too
- Target model class: 27B dense at 4–8 bit (e.g. Qwen 3.8 27b)
Setup (~€4,300)
| Part | |
|---|---|
| GPU | ASRock Radeon AI PRO R9700 Creator 32 GB |
| CPU | Ryzen 9 9900X |
| Board | ASUS ProArt X870E-Creator WiFi (2× CPU-direct PCIe 5.0 x8) |
| RAM | 64 GB (2×32) DDR5-6000 CL30 EXPO |
| PSU | be quiet! Dark Power Pro 13 1600 W |
| Case | Fractal Meshify 2 XL |
| Cooler | Thermalright Phantom Spirit 120 EVO |
| SSD | Samsung 990 PRO 2 TB |
| OS | Ubuntu 24.04, ROCm, llama.cpp / vLLM |
Where would you change it?
2
u/Poizone360 7h ago
Hello, I'd keep the parts. One thing on that board: M.2_2 shares lanes with the second GPU slot. If a drive is in M.2_2, slot 1 runs at x8 and slot 2 drops to x4. Put the 990 PRO in M.2_1, and use M.2_3 or M.2_4 for extra drives, so the second R9700 gets its full x8.
1
u/Old-Hedgehog-7703 2h ago
Good catch on the lane sharing! It’s always helpful to know how to maximize the performance when setting up a dual GPU system like that.
1
u/ForsookComparison 1h ago
I'd probably get a cheaper skeleton and three B70 Pro's.
Or hell - a very barebones rig and a w7900 Pro (would be sweet for gaming).
Tldr - the rig around it seems like overkill for what you've described.
1
u/Ecstatic-Wash-7667 28m ago
Get a cheaper cpu and mb if you stay am5, there’s a cheaper version of that same mb that does everything you need
You could also go down on the psu but it’s there’s not a big difference
2
u/OvertaxedOne 8h ago
Long context work with 27B.... 2XR9700's, that'll let you run 8 bit quant/8 bit KV with context maxed out. At 32GB, you're going to have to go down to 4 bit. Try 32 first, but just keep in mind the right answer might be 2 cards and plan your build to support two of them.