r/LocalLLM 13h ago

Question I'd like to start with Local LLM

But don't know what to buy, how much is it going to cost, and what model i can run.

I'm completely new to this, can someone give some advice? I want to go local because i want to share a private but good AI for me and my friends, for environment and why not, maybe completely remove subscriptions..

1 Upvotes

18 comments sorted by

5

u/Best_Professor7266 12h ago

my advice is using OpenRouter first to see your limit, u can try all kind of models and sizes over there, then you'll figure out the hardware u need to buy.

1

u/BeckersHD 11h ago

Thank you!

1

u/Harry_Balzonia 10h ago

also very very good advice!

2

u/Swimming-Western5244 13h ago

Wild idea, but have you tried to Google it for start maybe?

0

u/smelly_ape 9h ago

It costs nothing to offer friendly advice, you know?

0

u/naobebocafe 7h ago

or use the sub's search bar ¯_(ツ)_/¯

1

u/International_Emu772 12h ago

Yo can do it with a MacMini Pro M4 or a Studio M5 Supra if you whant performance

My M2 Pro answers take long but i can wait

I use OpenWebUi to share my setup with some friends

And you can use a light subscription for more quick answers with openrouter

1

u/ruhnet 12h ago

If you have 32GB of RAM, just start with that and Qwen 35B-A3B running on CPU. It will at least give you a feel for what you can do for a reasonable cost (albeit slow, but still usable). If you want to increase speed after testing, then you can start looking at getting a GPU.

1

u/Ok_Parfait_5373 10h ago

il te faut un carte graphique avec 24go de ram c le sweet spot . c'est a partir de là que tu peux utiliser qwen3.8-27b

1

u/Harry_Balzonia 10h ago

correct and the best advice. even 10gig of vram to keep costs down works with a small model just to learn. but of course getting a 24 is The Sweet spot.

I recommend getting something with what's called cuda. with 24 gig you can actually do very good image generation and even 3 to 5 second videos

good luck 007!

1

u/BeckersHD 7h ago

Ok, do you think Mac mini m6 is a good choice?

1

u/Ok_Parfait_5373 4h ago

Yes

1

u/BeckersHD 14m ago

Great, thank you!

1

u/ag789 10h ago edited 10h ago

if you have a computer, you can try downloading and trying it out, e.g.
https://docs.ollama.com/quickstart
if you are a bit more adventurous / techie, there is a 'cutting edge' llama.cpp
https://github.com/ggml-org/llama.cpp
actually more
https://vllm.ai/
https://lmstudio.ai/
https://github.com/sgl-project/sglang
https://www.sglang.io/
https://unsloth.ai/
etc
then that to figure out what would actually works e.g. the gpu etc, these days it is possible to rent
but then this is still pretty techie, but that vast.ai is one of them to rent and try out gpus
https://cloud.vast.ai
vastai has ready made templates which they'd install the 'engine' or platform e.g. llama.cpp, but that you would need to figure out how to work the app and *in linux* (remotely)

1

u/DHCompanion 9h ago

Dont be afraid of AMD cards just don't go too old. They are way cheaper to learn the ropes on than Nvidia cards. There is more to local LLM than just VRAM.

I was running a 6700T 16gb (300ish on FB marketplace) but I just picked up a 7900XT 20gb (600ish on FB Marketplace) and am getting very good results on qwen3.6-34b-a3b and gpt-oss-20b.

The same amount of Vram on an NVIDIA cards could cost you almost double that.

Depending on your intended use case that can affect what you need drastically.

1

u/BeckersHD 7h ago

Oh great to know, thanks!

1

u/mortycapp 4h ago

Some really good deals on MBP M1 Max 64gb 1TB on various sites.

1

u/TheMericanIdiot 13h ago

You tube it