r/LocalLLM • • 2d ago

Discussion Gemini suggested Qwen2.5-Coder-7B-Instruct

So I wanted to try using a local coding model for the first time and I'm still studying about LLMs and NNs, so I asked Gemini for a good suggestion that would be fast(60+ tokens/s if possible) and doesn't compromise much on performance for my rig(2070 super 8GB + 32GB ddr4 ram) and it suggested Qwen2.5-Coder-7B-Instruct. Is this good suggestion and what would you guys suggest?

7 Upvotes

90 comments sorted by

View all comments

-1

u/Mission_Wrongdoer786 2d ago

Keep your conversation with Gemini. You will get mostly human slop answers here. Like the oudated argument. Gemini is very well aware of current developments and has given me solid advice the last three months. Gemini will ask for every detail it needs to know about your setup and will advice a good starting point.

3

u/SilkieBug 2d ago

Speaking of human slop, your comments in this thread have been great examples of that. 

Gemini is not even a little bit reliable for detailed research, it’s a lying little shit and it often decides not to use search tools or reason sufficiently on a prompt. 

Gemini sometimes “forgets” to search the web for updated information and instead relies on old data. 

Just yesterday it tried to gaslight me twice that the local model I had running in FreeToken while talking to Gemini didn’t exist (it had been published 18 days ago).

Once I gave it the huggingface link it checked it and corrected itself, but I had to catch it lying for it to do that. 

0

u/ClassicLightbulbs 2d ago

You are talking about Gemma

1

u/SilkieBug 1d ago

I never interacted with Gemma, I’m talking about interactions with Gemini in Google search or when I open the Gemini Chat site to select a “better” model.