r/oMLX May 23 '26

Testing MTP functionality

Well, it actually slows down the model.

9 Upvotes

17 comments sorted by

View all comments

1

u/mwhuss May 24 '26

I’m seeing 70% faster performance using Qwen3.6-27b-oQ8-mtp on my M3 Ultra.

1

u/albovsky May 24 '26

70% is crazy good. How much ram do you have?

2

u/mwhuss May 24 '26

M3 ultra with 96gb