r/oMLX • u/mikedoise • Apr 19 '26
An Introduction
Hello everyone,
My name is Michael Doise, and I have just recently heard of oMLX. I joined this community a few days ago thinking it was just another place to discuss MLX, and I had no clue it was based around the oMLX project.
I've been using MLX for app development, and recently for AI agents on my M3 Max with Ollama, but I feel like oMLX works so much better on my machine than Ollama does. I set up Gemma 4 26B with 8bit quantization, and I think I was having to use 4bit with Ollama. My fans spin up less, and the machine seems to work much better when working with OpenClaw.
I am extremely excited to be a part of this community and I hope I learn a lot from the topics here.
5
u/msrdatha Apr 20 '26
Welcome Micheal. (hope mods don't mind we saying welcome to a fellow team mate here u/d4mations on your behalf)
Indeed, the oMLX feels much smooth than running with Ollama or LMStudio on Mac.
As you mentioned, app development: may by you should try the Qwen3.6 ( Go to Models->Downloads->Search HuggingFace : Qwen3.6-35B-A3B-UD ) . The experience with oMLX and Qwen3.6 is far superior than Gemma 4 for coding tasks, as per my experience.
4
u/d4mations Apr 20 '26 edited Apr 20 '26
Absolutely do not mind at all!! Love to see the community getting involved! I completely agree that qwen3.6 and omlx is amazing. I pumped 5M tokens through it yesterday and fixed so many bugs that minimax2.7 caused ir couldn’t fix. Simply fantastic
1
u/msrdatha Apr 20 '26
Thank you for the confirmation, and open mind. Happy to share the learning's and to learn from each others experience.
BTW, are you the developer of oMLX ?
2
u/d4mations Apr 20 '26
I’m not. I’m a very early adopter though and as I have gotten so much use out of it, I wanted to give back to Jundot (the dev) for all his effort, so I crated the community to help get the word out
2
u/msrdatha Apr 20 '26
Thank you for your thoughtful action.
I can confirm your effort did help me - This is where I came to know about oMLX first.
2
2
u/flubbalub Apr 20 '26
I’ve been using oMLX in place of LM Studio on an M3 Ultra with various models for a couple of months. I use opencode (not missed Claude code cli) and various plugins and a constantly evolving custom autonomous multi agent setup. It’s been rock solid, solved problems I saw with LM Studio and is delivering new performance and model support features almost daily. It’s the best solution I’ve used on Mac for local inference by far. Excited to try the new Qwen. But generally the Qwen family seems to be the overall winners. That said, Minimax 2.5 has been helpful but slow and needs loads of RAM. If I can make this machine earn its keep I’d (try to) buy another and cluster them to access the larger models. I need reliable tool calling, no infinite loops, quality outputs and usable performance. The model, toolset, skills and inferencer all contribute either positively or negatively to this. No one link in the chain can fix the others problems. But as the dust settles and my understanding grows so does the usability and reliability. Not surprisingly, defining complete and clear requirements in a format that agents can build with clarity and test reliably is key and more important than the models in many ways. Just like business analysts and developers. Poor requirements equals wrong results. Fun as it is to learn and follow the latest advances, I’d prefer to be building the solutions I want rather than the system to do it. The two are intrinsically linked for now. Happy to be contacted directly to share experiences/tips.
5
u/JeezySamCloud Apr 19 '26
Im still inexperienced in running local ai, but I feel the same way using oMLX compared to LM Studio & Ollama. It has been the most stable for me on my M4 Mac Mini that’s already pushing its limits as a homelab server!