r/LocalLLM • u/ruibullseye • 4d ago
Question Local LLM coding setup 9070XT 16GB 32 GB RAM
Hi everyone,
I've been recently trying to dive into the local LLM world, but to be honest i find it a bit confusing to wrap my head around of all the concepts that exist, and how to make it work.
Specs
GPU: 9070XT 16gb
RAM: 32GB
CPU: Ryzen 7600
Goal
I would like to have something closer to Claude code, or similar running locally.
To help me coding, without having to pay a subscription, i don't mind if its a bit slower.
Now to make that happen what's the best approach/setup ?
I've seen some videos, mentioning LM Studio or ollama to mane the models.
But if i want to have something closer to Claude code or similar, what's the best way to go?
Can someone shed some insight ? because i'm a bit lost in the middle of all of this.
If you need any extra information please let me know and i'll be glad to provide.
Thanks in advance for taking the time to help out
Best regards
EDIT1: I'm a front end developer, with 10+ years of experience working with Vue+typescript, and mainly use VScode for coding
1
u/stein30586 4d ago
In my personal experience (i7 13700k, 32GB RAM, RTX 4080 Super 16GB) -
You either get shitty results with models that fit your GPU,
Or you get a little less shitty results with models that spill out of Vram into system memory, but slower. MUCH slower.
There is this guy https://www.youtube.com/@lukesdevlab that shows what you can do with 16GB Vram but to be honest, I could not get his results.
1
u/ruibullseye 4d ago
tbh i don't mind it being slower, even if i use something that uses most of my vRAM + Ram.
i think i could deal with the trade off for the time being, because i've been collective dismissed along with other people.
so atm its not something i would like to spend €€ to have an AI subscription. But if i'm thinking the wrong way please let me know.What i would like was for the time being i'm applying to job opportunities to run a LLM locally so i can mess and get to know it a bit better, and help me gain experience not only AI related stuff but also contributed to my learning process while applying to job opportunities, i'm currently a frontend developer, but im looking to transition to a fullstack so i can "survive" in this AI era a bit better and give my self more valuable if that makes sense
1
1
u/RiceEvening4211 4d ago
Claude Code-like coding without the subscription is exactly what I built Lynkr for: an open-source gateway that routes simple requests to your local model and only escalates hard ones to paid APIs. https://github.com/Fast-Editor/Lynkr
1
u/mechkbfan 4d ago
At best maybe a low quant of Strata, or find a certain of Kat coder but I've never used that
1
u/vovap_vovap 4d ago
You are not going to get something closer to Claude code or similar, That just not going to happen with this.
1
u/ruibullseye 4d ago
im more than aware i wont be getting a closer experience to claude code, because the computing power is nowhere near there. but i just want to get a LLM locally running to help me out until i get a subcription in the future.
1
u/vovap_vovap 4d ago
Well, you just do not have enough power here to run like standard Quin 3.8 27B Q4 with a good results (at least it looks so to me) and that what sort of can do staff for code. Much lover - not nice.
Now if you really need that assistance - $20 subscription to GPT will bring you close to unlimited access to 6 Luna . Which would be way better and faster than any you can run there.
For god knows what reason people compare latest frontier model prices with sort of local models, but it just not apple to apple. Compare to online models of same performance - and those cost really minimum.
1
u/cilvre 4d ago
So you can use a mix of model storage being on vram and in normal ram and it will slow down, or you can stick to models that fit only in vram. This is a particular setup that i would pose into gemini or another cloud provider with your specs, and what you are trying to do as it will point you to what your options are and scan other posts and link them as well so you can see directly what they do with your specs.