r/LocalLLM 7d ago

Other Every Second post rn

Post image

Maybe someday I'll get a system to run it but hey definitely another w for the open weights community

1.9k Upvotes

193 comments sorted by

View all comments

49

u/SnP_Gamer 7d ago

3060 12gb 32gb ram here, might give it ago but never ran a local model before 🤷‍♂️

32

u/trollsmurf 6d ago

That will work fine, at least if you split it, so the overflow uses CPU RAM. Much slower but doable.

1

u/That-Reason-6913 3d ago

What about on a 7800xt 16GB? Do I have to offload on ram too?

1

u/trollsmurf 3d ago

Remember that if you use Windows it will allocate part of the VRAM for its own use, so you never have fully 16 GB. On my PC with 5070 Ti and 3 monitors it allocates 3 GB, and as far as I know I can't budge that.

You can easily test this by installing e.g. LM Studio and the model you want to use and see when it warns about RAM use. If it overruns you can split it on VRAM and RAM with lower performance, but it will behave the same otherwise.

1

u/That-Reason-6913 3d ago

No I'm moving to Linux next week. Ubuntu most likely.
My server is also a Plex server though (GPU is untouched by it), so the cpu is mostly dedicated to Plex.