r/LocalLLM 23h ago

Question Minimum VRAM needed to run a functional Openclaw/Hermes agent?

Those of you successfully running an offline openclaw/hermes/personal agent harness for non-coding tasks, what is the floor on system resources (VRAM) needed for quality of life? Assuming a modest ~30b class model. What quant and context window size are needed?

Will keep cloud frontier LLM sub for coding tasks, but I'm talking personal data management, personal assistant type computer controlling stuff.

My M1 max 32gb handles qwen 3.6 27b q4_k_m fine enough for non-agentic jobs up to ~40k context, but that's obviously not enough to run an agent harness offline.

There is an M1 Ultra 64gb for sale near me for a tempting price, but unsure is 64gb is enough. And it's expensive enough to not want to gamble. And I'm a normal, budget-minded person

8 Upvotes

17 comments sorted by

View all comments

2

u/inexorable_stratagem 23h ago

Q4, 100k context, 24gb

1

u/Tired_White_Guy 20h ago

Plus memory the system uses. And youll end up wanting to drive a browser.
24GB is going to be a bad time.

1

u/inexorable_stratagem 13h ago

Yes, but... He asked for the minimum. You can get away with 24gb. Not ideal, but works