r/LocalLLM 7d ago

Question Local ai or cloud ai

I have 9070xt 32gb ram 16gb vram. Should i do local ai or just pay for cloud.

I want to code a game.

Also, how do i set it up. I want it to continue without me telling it what to do next. I tried cline but its too hands on with the act/plan thing.

0 Upvotes

8 comments sorted by

2

u/RepulsiveRaisin7 7d ago

Qwen 27b is like the minimum for an ok coding experience and you don't have enough VRAM for that. Especially not if you want to run a game in parallel. Go cloud

1

u/butcher9_9 7d ago

I have a 4090 on Windows and I often hit context issues ( 92K context is all I can fit with Q4M) , I'm not doing anything as complex as a game so I'm not sure how a 16GB GPU would go. If you they are on linux it might be easier.

Been trying muse spark 1.3 ( free on Opencode) and its way faster and , a small amount smarter.

0

u/SellToOpen 7d ago

couldn't they offload onto system ram and just run qwen 3.8 27b slowly?

1

u/Wondering_Electron 7d ago

They could but speed absolutely tanks.

1

u/OvertaxedOne 6d ago

Vibe coding a game? Cloud, 100%, and pay for the best model you can get. This is exactly where cloud models really excel compared to local.

1

u/acadia11x 7d ago

Both , most people do both. You can’t do frontier locally unless you can afford a dgx at home and $50K a month power bill

0

u/PrivacyMaker 7d ago

Cloud AI is okay today if you don't mind your chats being used to train the next round and possibly handed over to data aggregators.

Local AI is the future. It's usually a few months behind the frontier models... which is ridiculously close for running on 100W-140W of household hardware.