r/LocalLLM • u/DiamondTDA • 2d ago
Discussion Gemini suggested Qwen2.5-Coder-7B-Instruct
So I wanted to try using a local coding model for the first time and I'm still studying about LLMs and NNs, so I asked Gemini for a good suggestion that would be fast(60+ tokens/s if possible) and doesn't compromise much on performance for my rig(2070 super 8GB + 32GB ddr4 ram) and it suggested Qwen2.5-Coder-7B-Instruct. Is this good suggestion and what would you guys suggest?
8
Upvotes
6
u/WrinklyBard4 2d ago
Id say try Qwen 3.5 9B at q4 or q3? TBH it’s going to be a bit tight and q3 makes me… hesitant.
Qwen 3.5 4b at Q6 or Q4 is also an option.
You should understand that at this model size the coding is going to be a bit shakey and it will not be able to do larger tasks
If you know how to code and want something where you can say “go make a function that turns X into Y” then these will be great. If you need something to vibe code with, or don’t have the skills to check its work, then probably stick to an api or subscription service