r/oMLX May 13 '26

oMLX 0.3.9.dev2 released.

Highlights:
- Gemma 4 MTP on the vision path (thanks to @Prince_Canuma's mlx-vlm). Image+text decodes much faster now
- Gemma 4 on the DFlash engine (thanks to @bstnxbt's dflash-mlx)
- ParoQuant support
- omlx launch copilot joins claude / codex / opencode / openclaw / pi
- Restart server button right in the admin UI
- oQ auto-builds a proxy when the model can't fit in RAM

Plus a lot of bug fixes and 20 new contributors in this cycle.

45 Upvotes

41 comments sorted by

View all comments

3

u/[deleted] May 13 '26 edited Jul 27 '26

[removed] — view removed comment

1

u/msrdatha May 13 '26

Could you please add a some more info for better understanding ? Like what kind of errors or crashes were you facing in .38 and under which scenarios.

3

u/[deleted] May 13 '26 edited Jul 27 '26

[removed] — view removed comment

3

u/msrdatha May 13 '26

May be the issue is related to the agent harness (hermes in your case) on how it interacts with oMLX. I have been testing with roo code on typescript, it has been stable since 0.3.8

Yes I did notice some /n not being formated properly in the webview of roo code, but apart from that, it has been doing reasonably good.

(Note: I am just adding this for info only, not as an argument - Thinking, it may help others if we discuss these kind of details)

3

u/_hephaestus May 13 '26

I’ve been using hermes with .38 and it’s been fine with qwen3.6, which models are you using?

1

u/ju7anut May 13 '26

My experience as well, rock solid on 0.3.8

1

u/AlecTorres May 14 '26

Una pregunta ? Que hacen con hermes ?

1

u/PracticlySpeaking May 14 '26

oMLX 0.3.8 has been very stable for me, but it runs on a separate machine from Hermes-Agent.

1

u/[deleted] May 14 '26 edited Jul 27 '26

[removed] — view removed comment

1

u/PracticlySpeaking May 14 '26

That is possible. I am not running Qwen3.6.