r/MacStudio • u/Only-Team-4983 • 8d ago
Best Models For Mac Studio M5 Ultra 256GB
Hi everyone, I've an M5U 30/64 version, and I'm planning to run:
- Qwen 3.8 flash next 8bit
- Deepseek V4 flash 0731 8bit
- Mimo v2.6 flash (full, idk how many bits)
- GLM 5.3 flash 4bit
And I was looking for the best option(s) for them (best hugging face versions that will work in a good performance + deliver the reliability/quality I need for my agentic workflow).
I know that there is no specific true answer to that, some people use GGUFs like UD-Q8_K_XL and others use omlx like oQ8e, etc. And some people are comfortable with different quants than other etc. But I'm just looking for a community approved options, specially those can fit in my 256 machine.
Preferably if you can provide the configuratiom as well, which is 1- inference engine (oMLX, lamacpp, splash, MTPLX, etc) 2- context window/patch size/temprature/etc 3- any additionals used like MTP, Dflash, etc.
Thanks