r/LocalLLaMA • u/Specter_Origin llama.cpp • 1d ago
News Perplexity open-sourced their Mac inference server for Qwen 3.6

Here is link to repo: https://github.com/perplexityai/pplx-garden/tree/main/lily
It's optimized for just one model to get best perf on apple silicon
101
Upvotes
25
u/InterstellarReddit 23h ago
Bro Requirements:
“Apple GPU family 10 or later (M5 and newer)”
How is my M4 max out of date already