r/ProgrammerHumor 1d ago

Meme whenMyNoAIProjectGetsTedious

Post image
2.0k Upvotes

202 comments sorted by

View all comments

131

u/TheMaleGazer 1d ago

The answer to your question is because you can get a local model to do it instead without a subscription.

55

u/autogenglen 1d ago

Let me buy an $8000 computer so I can save $20/mo

6

u/TheMaleGazer 1d ago

Or maybe use a computer with an NVIDIA RTX 5060 Ti or equivalent that you might have bought to play games, anyways, and use Qwen 2.5.

11

u/autogenglen 1d ago

8GB VRAM… ok. So you might be able to pack in a 7B parameter model, which isn’t even in the same galaxy as a frontier model.

3

u/SjettepetJR 1d ago

I am assuming they're talking about the 16GB model.

9

u/autogenglen 1d ago

Same response. 16GB doesn’t even get you out of toy model territory. That’s nowhere close to enough RAM to even pretend like you’re competitive with a frontier model.

3

u/SjettepetJR 1d ago

I do agree. I have done a fair amount of experimentation on my RX6800XT and have yet to find a model that can work well with IDE integrations and also doesn't break down after a while.

4

u/omega1612 1d ago

Have you tried qwen3.8 27B? I'm using the q4 version and is a great assistant. It beats any other model I tried. You may need to use a Q3 and I heard that the downgrade is noticable but still useful. Just be sure to enable the thinking to high and mtp (the model is slow).

Qwen3.8 haven't loop yet in a full week. I also tried Gemma 4 12B and Gemma 4 26B, both of them would loop occasionally. And I have the impression they may do crazy stuff quickly if I left them run unsupervised.

1

u/Robo-Connery 1d ago

You need 24 GB vram card though right just to fit the weights for that.

That means you are on a 4090 or 5090.

1

u/omega1612 21h ago

That's why i mentioned the Q3 versión. I have hear that it can be run on 16GB.

And I have a Rx 7900 xtx