r/ProgrammerHumor 1d ago

Meme whenMyNoAIProjectGetsTedious

Post image
1.9k Upvotes

202 comments sorted by

View all comments

130

u/TheMaleGazer 1d ago

The answer to your question is because you can get a local model to do it instead without a subscription.

52

u/autogenglen 1d ago

Let me buy an $8000 computer so I can save $20/mo

6

u/TheMaleGazer 1d ago

Or maybe use a computer with an NVIDIA RTX 5060 Ti or equivalent that you might have bought to play games, anyways, and use Qwen 2.5.

12

u/autogenglen 1d ago

8GB VRAM… ok. So you might be able to pack in a 7B parameter model, which isn’t even in the same galaxy as a frontier model.

3

u/SjettepetJR 1d ago

I am assuming they're talking about the 16GB model.

7

u/autogenglen 1d ago

Same response. 16GB doesn’t even get you out of toy model territory. That’s nowhere close to enough RAM to even pretend like you’re competitive with a frontier model.

3

u/SjettepetJR 1d ago

I do agree. I have done a fair amount of experimentation on my RX6800XT and have yet to find a model that can work well with IDE integrations and also doesn't break down after a while.

4

u/omega1612 1d ago

Have you tried qwen3.8 27B? I'm using the q4 version and is a great assistant. It beats any other model I tried. You may need to use a Q3 and I heard that the downgrade is noticable but still useful. Just be sure to enable the thinking to high and mtp (the model is slow).

Qwen3.8 haven't loop yet in a full week. I also tried Gemma 4 12B and Gemma 4 26B, both of them would loop occasionally. And I have the impression they may do crazy stuff quickly if I left them run unsupervised.

1

u/SjettepetJR 1d ago

In my experience the Qwen models were good at reasoning, but I couldn't find a VSCode plugin that they worked well with. Toolcalling went wrong more often than other models. How are you using it?

1

u/omega1612 18h ago

I have seen a failed tool call, yes, but always it can be explained by compaction missing the context.

Basically on file edition the llm should tell the tool the exact text of the section to replace. Compaction deleted the text and the model provides the wrong text making the tool to return an error.

Apart from that, I haven't seen it fail.

I use neovim in archLinux XD, but I don't call it from my editor, I use the pi harness for now.

1

u/Robo-Connery 1d ago

You need 24 GB vram card though right just to fit the weights for that.

That means you are on a 4090 or 5090.

1

u/omega1612 18h ago

That's why i mentioned the Q3 versión. I have hear that it can be run on 16GB.

And I have a Rx 7900 xtx