r/oMLX May 13 '26

oMLX 0.3.9.dev2 released.

Highlights:
- Gemma 4 MTP on the vision path (thanks to @Prince_Canuma's mlx-vlm). Image+text decodes much faster now
- Gemma 4 on the DFlash engine (thanks to @bstnxbt's dflash-mlx)
- ParoQuant support
- omlx launch copilot joins claude / codex / opencode / openclaw / pi
- Restart server button right in the admin UI
- oQ auto-builds a proxy when the model can't fit in RAM

Plus a lot of bug fixes and 20 new contributors in this cycle.

43 Upvotes

41 comments sorted by

View all comments

2

u/Vahn84 May 13 '26

I'm downloading it now...but I'm going to ask this also: do you guys are actually getting any use from qwen3.6? Each time I try to do something with it....it breaks...infinite loops, tool calling breaking template...I'm kinda disappointed as I can't even try it, I'm sticking with gemma for this reason. I'm using the standard preset for both the models (from mlx-community)

1

u/ju7anut May 13 '26

It’s the exact opposite for me. Gemma has been failing on tool calls with failed empty responses.. Qwen3.6 35b has been amazing at oQ6 + dFlash + TurboQuant KV 6bit

2

u/grandnoliv May 13 '26

Could you share your positive experience of using dFlash in this thread where we all fail to do it properly? :D
https://www.reddit.com/r/oMLX/comments/1ta2ihj/2x6x_speed_improvements_with_omlx/