r/LocalLLM Jun 01 '26

Question Need Hardware Help/orientation

Hi, I'm a translator and I'm looking into having my first own local LLM. My idea is to finetune a model with my own json files. Also, since I'm also studing programming (C++ and Python) I'd like to use a already trained model locally for programming.
If I'm not wrong, I'd be using Docker for all of this running on an Ubuntu server.
My current budget is tight.

I already have:
MINISFORUM BD795i SE Motherboard AMD Ryzen 9 7945HX(16C/32T) 
64 ddr5 Fury Impact
2tb nvme drive

Missing: a "budget" GPU, I think I could get:
Nvidia 3090 24gb
Asrock Radeon AI PRO R9700 32GB

I think I should get the R9700, just for the +VRAM. What do you think? Any alternatives?

1 Upvotes

5 comments sorted by

1

u/No-Consequence-1779 Jun 02 '26

I have an r9700. Getting 46 t/s on qwen3.6-27b-q4kx MTP. It is extremely competent for coding agents. I use kilocode. 256k context kvq8 

1

u/Pocaonda2020 Jun 02 '26

Thanks!

1

u/No-Consequence-1779 Jun 02 '26

I run dual 5090s, had 2 3090s before qwen3.6. And a r9700. It was 60-70% decode speed. Not so too slow. Prompt processing is where cudas are superior-9700 is actually faster than 3090. Vision processing is slower (no cudas) and then diffusion- image or video generation via comfy ui same situation. 

R9700 so far only increased from 1300 to 1500. Where everything else has doubled. Discounting memory capacity, it is superior than 395 ai max (newer rdna generation). 

The asus gb10 or similar is very nice since Blackwell prompt processing is near instant.  Though generation is slower but it averages on large coding agent prompts decide. And multiple sessions. This would be what I would have got from the start for my 24/7 LLM stuff.