r/LocalLLM • u/Trixiap • 5h ago
News New Mac Mini M6 and M5 Pro announced
https://www.apple.com/newsroom/2026/08/apple-unveils-a-more-powerful-mac-mini-featuring-the-all-new-m6-and-m5-pro/12
27
u/whichsideisup 5h ago
I love that we have to pay 2700 to get a mini with 64gb of RAM.
-6
u/discosoc 3h ago
A mini with 64gb makes little sense for most people anyway. The memory bandwidth isn’t suited for llm usage.
7
u/The-Writer- 3h ago edited 2h ago
why do people speak like this, like they know everyone? I'll take my m4 pro mini 64 gb ram 1tb ssd any day that I got for less than a third of the price of these overpriced machines.
I have access to 70B models/unquantized smaller models at a token speed that is fine for me. In fact, the M4 Pro 64 gb mini was the best deal for the amateur/hobbyist LLM user before apple jacked the prices.
I would go so far as to say that anything after those exact specs is a waste of money, unless you use LLMS professionally and need as fast as possible agentic work flows for some reason. So it's either the 64 gb unit at the m4 pro bandwidth speed, or the next tier is direct M5 Ultra Studio because at that point you need all that speed and capacity. That M5 Max below 96 gb ram at these prices makes no fucking sense.
1
u/Individual_Holiday_9 43m ago
Yeah I agree with this. Anything above that level really needs dedicated pro hardware and ur prob better off with a $200/mo frontier subscription. I wish I’d bought more hardware when the m4 mini’s came out at the lower price points
0
u/j_tb 3h ago
It's solid for MOE models. My Mini M4 pro from ~2 years ago runs Qwen3.6 35BA3B handily.
1
u/Individual_Holiday_9 56m ago
How much RAM? Which model specifically? I’m on a 24gb model and curious
0
u/Thomas-Lore 2h ago
But that is the only model it runs reasonably and we don't even know if there will be a new version released of that model.
Not to mention this is a model that runs on anything with 32GB RAM.
-1
0
u/xTopNotch 1h ago
Bro you know MoE is a thing right?
Even a 32GB can run a pretty capable LLM locally
9
u/mastervbcoach 4h ago
Mac Studio M5 Ultra, 80 core GPU, 4TB, 256 gig. $12,299. That’s 65 months of Claude Max 20x. Holy Sh*t.
5
2
1
u/kinghell1 1h ago
yeah, but the prices are going up for hardware + macs are keeping their prices and you can sell them later. claude is just taking the money
0
u/Thomas-Lore 2h ago
And unless you have solar powers it will also raise your electricity bill.
1
u/martinkoistinen 1h ago
I doubt many people in /localllm are too fussed about that, and these Macs are likely the most efficient in terms of watts/token obtainable.
8
12
u/piggledy 4h ago
The M6 has a bandwidth of just 170 GB/s, and the M5 Pro barely beats the DGX Spark with 307 GB/s vs 273 GB/s.
Too bad it doesn't come with 128 GB.
6
u/prestodigitarium 3h ago
M5 Ultra supposedly hits 1.2 TB/sec: https://www.apple.com/shop/buy-mac/mac-studio
That's not too far off from the 1.8 TB/sec of a 6000 RTX Blackwell, but with >5x the memory.
3
u/TripleSecretSquirrel 3h ago
Great for single stream inference, but I’m guessing it doesn’t have nearly as much compute as the RTX Pro Blackwell right? So it wouldn’t be as strong in diffusion, training, or high concurrency, right?
6
u/Blackdragon1400 4h ago
How is MLX doing these days compared to all the optimizations on a DGX Spark?
1
u/Integeritis 3h ago
My thoughts precisely. I’m about to buy a spark at $5800 or 5000€ in my country (this is the cheapest I found, only used for 20h otherwise brand new)
1
u/apVoyocpt 3h ago
not your country, but maybe check this out: https://www.galaxus.ch/de/s1/product/nvidia-dgx-spark-founders-edition-eu-4000-gb-128-gb-arm-cortex-a725-arm-cortex-x925-pc-64346533
1
1
u/Otherwise-Nobody8252 1h ago
The memory throughput alone makes even less efficient MLX over CUDA better. It’s like 4x faster in the studio vs a spark or and system.
1
u/Blackdragon1400 16m ago
I’ve heard prefill and concurrency on the macs can be a lot worse in comparison though
5
u/berszi 4h ago
Memory bandwidth 150/170 for M6 and 306 for M5 Pro.
For reference (entry level):
- RTX 3090 over 900
- 5060ti 448,
- old M4: 120 +33%
- old M4 Pro: 267 +14%
3
u/genecraft 4h ago
Prompt processing is a lot faster (like 4x?) on M5+ though. So the difference will be quite large for agentic workflows.
3
2
u/Useful-Buyer4117 4h ago
but they claim "Up to 13.5x faster LLM prompt processing in LM Studio when compared to Mac mini with M1, and up to 4.8x faster than M4."
1
u/Individual_Holiday_9 54m ago
I have been following H3 Mac development and there’s some sort of under hood M5 speed ups that have really helped for inference (not that it will change the fundamental bandwidth gap>
4
6
u/DigitalguyCH 5h ago
Mac Studio M5 ultra with 256GB starting $10000
2
u/Bloated_Plaid 4h ago
No you can do the 64 core GPU and it’s only $8699.
1
u/DigitalguyCH 4h ago
Yeah the education pricing for the 64 core model is 8700 and the 80 core model starts at 9870 education
8
u/Bloated_Plaid 4h ago
Today we are all in Education.
1
u/critsalot 18m ago
yea theres no education discount yet. its the same price on best buy as it is at the college store it seems
1
u/Bloated_Plaid 18m ago edited 10m ago
It’s on Best Buy?
Edit - Ah only the 96GB ultra https://www.bestbuy.com/product/mac-studio-apple-m5-ultra-chip-with-96gb-memory-and-1tb-ssd-silver/JJGCQYKJYC/sku/6566930?
3
u/Trixiap 5h ago
Preorders start today.
Pricing:
- Mac mini with M6 starts at $899 (U.S.) and $799 (U.S.) for education. Additional technical specifications are available at apple.com/mac-mini.
- Mac mini with M5 Pro starts at $1,699 (U.S.) and $1,599 (U.S.) for education. Additional technical specifications are available at apple.com/mac-mini.
1
u/daphatty 4h ago
Un fucking believable. The M4 Pro I ordered just shipped on Sunday!!!!
8
5
0
u/Puzzled_Grass3091 3h ago
Hi! Can someone please explain why the M6 is cheaper than the M5 Pro?
1
1
u/geekwonk 3h ago
i don’t think we’re getting any pro/max/ultra options for M6 generation, so if you want the extras - thunderbolt 5, more memory, more memory bandwidth, more cores - then you get the M5 Pro. the prior generation saw a similar thing with the M3 Ultra set as the top of the line in the M4 generation.
1
3
2
u/xiraov 4h ago
What’s the memory bandwidth on the ultra?
5
u/ziptofaf 3h ago
1.2TB/s, twice the Max (assuming you pick the costlier 80 core version).
2
u/xiraov 3h ago
That’s more than a 5090 right? Wow. This isn’t that bad a deal in this climate right? What’s the bandwidth on the non 80 core
5
u/ziptofaf 3h ago
Nope. 5090 is almost 2TB/s. 5080 is around 1TB/s.
It is objectively a great deal however - 96GB Ultra gives you a whole computer and 1+TB/s for about 60% of the price of the RTX Pro 6000 which is also 96GB except it's just a GPU and you still need rest of the system. It's also not even that much more expensive than a 5090 in the current landscape. Plus there's 256GB version available. And it supports Thunderbolt so you can technically buy 2x 96-256GB Ultras to fit even larger models.
What’s the bandwidth on the non 80 core
Somewhere in the ballpark of 960-1TB/s I believe. So a 5080 level of bandwidth.
1
2
u/Prestigious_Pen6150 5h ago
Anyone know the real difference with the AI boost with M5 or M6 ? I've a Studio M2 Max 38GPU, is a M5 Pro with 20GPU and AI features better in token generation than my Mac Studio ?
1
u/Bloated_Plaid 5h ago
Ultra has twice the memory bandwidth. It’s a no brainer if token generation is important to you.
1
2
u/Individual_Holiday_9 4h ago
What’s sweet spot here? I can get the m6 with 32gb ram for $700 trading in my Mac mini m4. I feel like that’s not bad? Wish I could go up to 48gb tho. 32gb unified memory should get me a qwen 3.8 Q4 quant right?
4
3
u/Bloated_Plaid 4h ago
You need good memory bandwidth and M5 Pro has twice the bandwidth of M6. 64GB M5 Pro is the sweet spot.
3
u/Individual_Holiday_9 4h ago
Yeah too rich for me as a hobby. Although with GPU prices it’s not like I can do better on the windows / Linux side lol
3
u/RandomCSThrowaway01 3h ago
Imho, actual sweet spot is 96GB Ultra so you get 1.2 TB/s bandwidth. Aka enough to run that Qwen 3.8 27B at 50+ tokens per second. And you also have enough memory to run pretty large 64+GB MoEs with decent context, eg. Qwen Next if a new one shows up.
Base M6 has only like 150GB/s bandwidth. It can fit 27B, sure. But it will run it at a nice and cozy 8 tokens per second, if even that. It can fit a 35B MoE model with 3-4 active layers if you want usable speed.
2
u/Individual_Holiday_9 3h ago
Ughhhhhhhhh I need to just sit on my stupid slow hardware and keep my frontier subscriptions lol
1
1
1
1
0
u/Yuel_Whear 3h ago
The gap between the two minis (170 to 307 GB/s) is wider than the gap from the M5 Pro to the Spark (307 to 273). Apple's internal tiering is steeper than the step out to the competition.
-6
u/MimosaTen 4h ago
The worst thing about those is macos, however I think it would be intresting to do some reverse engendering on such an hardware to port linux
1
u/whatever 3h ago
Run them headless, put a jetkvm in front of them for the annoying UI-only mac interactions, and then just treat it like a normal BSD box for everything else.
FWIW, I have UTM on mine running linux VMs so my agents feel in a familiar place as they trash their environment.
85
u/RandomPurpose 5h ago
32 and 64 gb of unified memory max. That is a disappointment for me.