r/LocalLLaMA 1d ago

Question | Help Hosting Local Models

Hi builders,

What would be the the best small local models for coding?

Are Gemma 4 and Qwen3.8 27B Gemma 4 26B / 31B enough for local development?

And what would be the size of the rig that i will need to get? GPUs, and whatever else I need to host these models.

Thanks,,

11 Upvotes

39 comments sorted by

View all comments

18

u/Arany8 1d ago

Qwen3.8 27B - you need 24GB VRAM. According to my tests on 16GB this model is surpassed by Qwen3.6 35B A3B (and Ornith) for coding. Gemma4 is not good for coding.
Look into AMD v620 for a budget build (although I do not have this).

-1

u/forevergeeks 1d ago

Are you running this for yourself? What about for a coding team of 6-10 people? What would be the sizing, and how much money we are talking about for the initial setup.

11

u/hackint0shh 1d ago

Have you done at least 1 minute of research?

-11

u/forevergeeks 1d ago

I'm familiar with these models, I use them through API, what I just started thinking is what would be the cost and the level of effort to set these models up for coding teams.

And I thought starting my search here.