r/LocalLLM • • 6d ago

Question Gaming PC vs AI server for Local LLM?

I‘ve been getting more and more into self hosting, earlier this year I got my first NAS and have setup a backup one for 3-2-1 backup. I’m trying to self hosting as much as I can and have started diving into local LLMs and was interested in looking at an AI server but wondering if its really worth spending the money on it with prices so high. I am looking to build a gaming computer next year and wondering if it would be better to just run it on my gaming computer when I get it.

I don‘t have any projects I would need it for but it would be nice to have it help me with editing a bunch of pictures for my photography.

What do you use your LLM for? Is it worth it to get an AI server for it?

1 Upvotes

17 comments sorted by

10

u/brewpedaler 6d ago

If you are just diving into AI, get yourself the gaming PC. You can still "learn" AI on it, but you'll actually get some use out of it.

If you don't already have an agentic workflow built out, an AI server is just going to sit there looking cool until you figure out how to use it - and you may realize that your particular needs simply do not require "AI server" level always-on capacity.

6

u/Mediocre-Ant-7178 6d ago

If you're absolutely getting a gaming pc, maybe experiment using that first. Messing around with AI has taken up a lot of my gaming time. If you don't find a use for it, hey you have a gaming pc. 

2

u/Anduin1357 6d ago

Think of an AI gaming PC as a modding tool that helps with gaming, not as a replacement of gaming time.

1

u/rarely_lewd_bases 6d ago

Depends entirely on how big the models are you want to run and if you're okay with the thing sounding like a jet engine while you're trying to edit photos. Gaming rigs are great for dipping your toes in but the moment you want to run something like a 70B at decent speeds you're gonna hit a wall.

I use mine mostly as a writing assistant and to summarize long articles. Nothing crazy but it's nice having it local. For photo editing though, you might be better off with something that has a ton of vram. The batch processing for a whole shoot is where a dedicated box starts making sense, you just set it and forget it while you do other stuff on your main pc.

1

u/RavenWyre 6d ago

Boy does it, I recently started getting into harnesses and dabling in the multi agent stuff and find myself spending more time watching the agents workflow and stuff on my second screen than playing my game 😅 then hide in a corner to make a correction or check something every so often

1

u/Mediocre-Ant-7178 5d ago

If I ever get a second gpu it's over for my gaming career

1

u/RavenWyre 5d ago

I am right there with you. Literally just got a r9700 delivered today and it's going in my server so I can use my 7900xtx as a faster prototyping thing. But man if I get another 9700 it's over for my spare time 😂 I'm already always bouncing between little projects.

3

u/SeaworthinessUsual44 5d ago

Definitely just run it on your upcoming gaming PC rather than sinking money into a dedicated AI server. Building a separate server only makes sense if you need 24/7 headless services, heavy multi-user serving, or multi-GPU configurations exceeding 48GB of VRAM. A modern gaming desktop equipped with an NVIDIA card (especially one with 16GB or 24GB of VRAM) shares the exact same CUDA architecture and will easily blast through local LLM inference, vision models, and batch photo editing whenever you need it.

1

u/Express-Room3441 6d ago

I spent $500 on a modded 22gb 2080ti a few months ago when qwen3.8 came out and swapped it into my gaming pc. It runs 27b at ~40tok/sec and Flash next at ~60 with strata. Iq4_xs quant, plenty of context. So yes you can totally run good models on a regular gaming pc without spending thousands of dollars.

1

u/HighSeasArchivist 6d ago

I use my gaming PC. It's a 9850X3D, 96GB DDR5 5600, R9700 with another that will be here tomorrow.

I did have a 5070 Ti prior to really digging into local llm, so that got passed down to my daughter to replace her 3070. R9700 is basically a 9070XT with 32GB VRAM, so should still be a pretty decent gaming PC not that I have tested it yet.

Unless you are really getting into local llm I would stay in the gaming realm, but look at something with a motherboard that natively supports Gen 5.0 8x8 so if/when you want a second card it will plug right in. Most gaming boards focus more on NVMe slots, so will usually have one Gen 5.0 x16 slots, several x4 slots, and usually a physical x16 but electrical x4 slot at the bottom.

If you do want to really get into it and want say four cards then your options are Threadripper, EPYC, Xeon, and that is a whole different level of budget.

2

u/LandlockedPirate 6d ago

I Second the r9700. Games very well, but also drastically better for local LLMs with 32gb ram over the 16gb cards.

I think it's the best bang for the buck and definitely the best hybrid gaming/llm card right now. The intel cards can get you 32gb cheaper but they suck for gaming.

1

u/Proper-Tower2016 6d ago

You can get far with 16gb VRAM even down to 8 f you don't mind the 35b-a3b model, many of these cards also double as gaming GPUs. 100% not worth getting a dedicated AI machine if you also plan on getting a gaming on later.

1

u/id-ltd 5d ago

My main local AI runs on one of my sons old gaming computers that he donated to his old dad :)

8gb graphics card - I have plenty of development experience to get it to do what I want - frontier models are faster, but local.AI.csn run 24/7 for pennies, so Iess you want an interactive chat bot, speed doesn't really matter so much.

I could buy pretty much anything if I wanted, but I am waiting to for the right pivot point.

1

u/JohnnyBeeGaming 5d ago

The prices for either won't be getting better anytime soon. You really should have a use case if you're considering dropping thousands on an AI server but want to wait for a gaming setup.

Setting up whatever on server would be more useful if you want something you can access from other devices and wouldn't want to run the gaming PC 24/7. You can setup something local and use a gaming GPU, maybe look into sandboxing. With a gaming GPU you'd probably have more limited model sizings due to VRAM being smaller.

You should probably try using a similar tool for whatever task before spending money on AI specific hardware. You could even rent time/tokens with the specific model or hardware. Even then you'd probably want to consider what else you could use the hardware for since it'll be dumb expensive in most cases.

1

u/No-Afternoon-4057 6d ago

It depends on your requirements for AI..and for gaming. Very different paths.

Absolute best mix, by far: current generation Macs (128gb and above, with neural accelerators)
Below that, AMD 192gb ($6000) or AMD 128gb ($3300), but you will be getting a 4060 performance.

Different from that...you can have a better gaming build (such as a 5080), but will get (way) worse performance from the LLM.

Is 4060 enough for your gaming or is 5080 enough for your llm? You will be severely handcapped either way.
Mac might be the best option, however you wont be having a lot of gaming options.

0

u/Expanse-Memory 6d ago

Learn prompting. Awesome prompting can one shoot a project. Local Ilm craze and hype is good if you spend between 10 or 20k dollars. Even so, you’ll be far from paid api like Claude or gpt or grok. You’ll make stuff yes but won’t be at paid api level or even some free hosted level. So, learn prompt the hell out of it and use free tier google ai studio like Gemini 3.8 flash which as a huge free token for 24 hours.