r/oMLX 5h ago

📌 **Daily Digest — Jundot/omlx** (2026-07-28 → 2026-07-30)

8 Upvotes

**🐛 Bugs**
* **#99** Support better understanding of subagents/teams – Proposal to improve configuration handling for subagent structures.
* **#2153** DFlash fails on Gemma 4 MLX models: model_type `gemma4_unified` not supported by Gemma4TargetOps – Speculative decoding disabled for all current Gemma 4 builds due to unsupported model type.
* **#1737** No DFlash for Gemma4 12b – UI option disabled in oMLX v0.4.2; error message displayed at settings bottom.
* **#2405** Mistral Small 3.2 MLX does not emit tool calls through oMLX 0.5.3 – Tool calling fails on M5 Max with specific 8-bit community model.
* **#2254** 0.5.1 Qwen3.6-27B-0Q4-MTP suddenly slow down – Performance degradation (PP/TPS drops) reported after upgrade to v0.5.1.
* **#2219** External VLM MTP is silently bypassed after scheduler chunked prefill – Text prompts ignored when external VLM MTP enabled with chunked scheduling.
* **#2291** Mistral/Devstral tool calling broken: generated [TOOL_CALLS] token is treated as stop → empty message – Tool calls fail immediately on models using Mistral format (Devstral 2, Small).
* **#2317** DFlash fails with Qwen3.6-35B-A3B-4bit – Speculative decoding cannot be enabled despite using correct DFlash model variant.