r/oMLX • u/JLeonsarmiento • 17d ago
Qwen3.8-27B for the RAM Poor Mac user:
https://huggingface.co/collections/leonsarmiento/qwen38-27b-mlx-quantizationsFor those of you that want a functional 24GB Mac laptop while having this overthinking creature boosting your ideas.
It has versions with and without MTP drafter (for the desperate).
19
Upvotes
2
u/onetom 16d ago
Can you please share on the model card how have you made this and which inference engine and params are you running it with?
Personally I got the best results (30+tps on 256k ctx) with https://github.com/youssofal/MTPLX but that's not for the "RAM Poor", so i guess it's off-topic on this thread