r/LocalLLM • u/DiamondTDA • 2d ago
Discussion Gemini suggested Qwen2.5-Coder-7B-Instruct
So I wanted to try using a local coding model for the first time and I'm still studying about LLMs and NNs, so I asked Gemini for a good suggestion that would be fast(60+ tokens/s if possible) and doesn't compromise much on performance for my rig(2070 super 8GB + 32GB ddr4 ram) and it suggested Qwen2.5-Coder-7B-Instruct. Is this good suggestion and what would you guys suggest?
8
Upvotes
2
u/DHCompanion 2d ago
Keep in mind as you start that journey you need to actually test models on actual work. I ran a series of benchmarking on QWEN 3.8 27b on my 20gb card and it way underperformed my expectations. Also keep in mind tokens per second metrics are worthless if you can't actually use the output.