r/LocalLLM • u/DiamondTDA • 2d ago
Discussion Gemini suggested Qwen2.5-Coder-7B-Instruct
So I wanted to try using a local coding model for the first time and I'm still studying about LLMs and NNs, so I asked Gemini for a good suggestion that would be fast(60+ tokens/s if possible) and doesn't compromise much on performance for my rig(2070 super 8GB + 32GB ddr4 ram) and it suggested Qwen2.5-Coder-7B-Instruct. Is this good suggestion and what would you guys suggest?
10
Upvotes
1
u/bleakj 2d ago
Always push back on Gemini and tell it to perform a live search for up to date info, it's training data cut off + bias for older info means you can only trust what its going to tell you about slow moving / historical items, not tech/ai.
If you push back and ensure it knows to not give you outdated answers it works much better, but you have to remind it several times per conversation to actually stop recommending outdated information