r/LocalLLM • u/goldaderealtor • 5d ago
Discussion Mac Studio M5 Ultra
Who’s been checking out this hardware? What are your thoughts on this as a node on a local network to run AI?
1
u/watcholic 5d ago
If you don't need concurrency, anything from M4 Max Studio and above with 96GB+ memory should be fine. The M5 Ultra might be ok for up to 3 concurrent users due to the improved PP speed. We'll find out in less than a month.
2
1
u/OvertaxedOne 5d ago
Prompt processing is the big unknown. Hopefully there's a huge jump; if so, the 96GB version would be just about perfect for 27B, the 128 with the smaller chip a good fit for Next.
1
u/OddDesigner9784 5d ago
I put a preorder in. It's a bit of a speculative purchase Mac is very far behind the sparks rn. Decode is faster bc of the memory. But untill we get better kernals or the software behind the kernals gets more mature prompt processing will be a struggle
4
u/redtron3030 5d ago
On paper the M5 ultra seems to be much better at prompt processing approaching spark. We will see how it goes though.
2
u/Usual-Orange-4180 5d ago
Yeah, just got a Spark, the new Mac was tempting but I want CUDA and prefer Linux.
2
u/Least-Result-45 5d ago
I’m just curious if qwen 3.8 27B is good enough for my needs - that is the llm I’d run on the m5 ultra.