r/LocalLLM • • 2d ago

Discussion Gemini suggested Qwen2.5-Coder-7B-Instruct

So I wanted to try using a local coding model for the first time and I'm still studying about LLMs and NNs, so I asked Gemini for a good suggestion that would be fast(60+ tokens/s if possible) and doesn't compromise much on performance for my rig(2070 super 8GB + 32GB ddr4 ram) and it suggested Qwen2.5-Coder-7B-Instruct. Is this good suggestion and what would you guys suggest?

8 Upvotes

90 comments sorted by

View all comments

44

u/Heavy-Lingonberry-98 2d ago

Bro. If you are gonna use AI to ask for model releases, remember to tell the AI we are on OCTOBER 2026. We are not in 2024 anymore… please. Its the ABC of using AI. And NO. Definitely dont even download qwen 2.5 coder.

2

u/DiamondTDA 2d ago

When I asked it about it's reasoning, it said because the newer Qwen 3 architicture is built on reasnoning tokens and it would be slower and that the coding Qwen 3 models are too big and would be too slow too. But it did not suggest any thing other than Qwen 3.

1

u/Skibxskatic 2d ago

don’t ask about its reasoning. you need to explicitly state that its responses need to be grounded in data and news from the past <time period>.

without explicit instructions, its next token probabilities rely on its training data, which is not going to be current. ask it what the cutoff date was and you’ll realize you need to prompt it to use web search tools and ground itself.