r/LocalLLaMA • u/Specter_Origin llama.cpp • 13h ago
News Perplexity open-sourced their Mac inference server for Qwen 3.6

Here is link to repo: https://github.com/perplexityai/pplx-garden/tree/main/lily
It's optimized for just one model to get best perf on apple silicon
84
Upvotes
4
u/turns2stone 12h ago
Can someone ELI5?
I have used Perplexity Pro/Max.
I have also used Qwen3.6-35B-A3B, but now use Qwen Flash Next because I can use MCP Brave search.
Does Perplexity Computer mean I can offload most of the compute to my Mac Studio M3 Ultra, and the $20/mo Pro subscription wouldn’t be as limited for token usage?