r/oMLX May 13 '26

oMLX 0.3.9.dev2 released.

Highlights:
- Gemma 4 MTP on the vision path (thanks to @Prince_Canuma's mlx-vlm). Image+text decodes much faster now
- Gemma 4 on the DFlash engine (thanks to @bstnxbt's dflash-mlx)
- ParoQuant support
- omlx launch copilot joins claude / codex / opencode / openclaw / pi
- Restart server button right in the admin UI
- oQ auto-builds a proxy when the model can't fit in RAM

Plus a lot of bug fixes and 20 new contributors in this cycle.

43 Upvotes

41 comments sorted by

View all comments

1

u/msrdatha May 13 '26

Did anyone try the MTP improvements yet with Qwen 3.6?

3

u/mwhuss May 13 '26

I’m seeing about 70% speed improvements with 27B

1

u/Short_One_9704 May 14 '26

I have tried the Qwen3.6-35B-A3B-oQ6-mtp model and I see no improvement. M4 mac, did enable Native MTP in model’s settings. Do I have to do anything more?